Skip to main content

Model Catalog and Pricing

This page lists 249 online LaoZhang API models with input, output, cache-read, per-call, and tiered prices, token groups, and supported endpoints. The data was updated on 2026-09-02, 12:12 (UTC+8). Confirm final availability and charges in the signed-in console and call logs.

Online models249
Vendors13
Usage-based228
Per-call21
Confirming prices and next steps

Online models, cache-read prices, and tiered pricing rules are refreshed from current system configuration. Use the signed-in console to confirm account groups and current prices, and call logs to confirm actual charges.

Pricing notes

All amounts on this page are shown in US dollars. Usage-based input, output, and cache-read prices are per million tokens ($/1M). Per-call models are priced for each call. The table shows default list prices; account contracts or dedicated routes may use different prices.

Enterprise purchasing and discounts

Contact the site owner or support for enterprise purchasing, volume usage, contracts, invoices, discounts, or special payment arrangements. The documentation does not publish fixed discount thresholds or percentages. Final commercial terms, credited balance, and charges follow the confirmed arrangement, console, and call logs.

A model can support multiple token groups with different routes, permissions, or billing modes. Confirm the groups and actual prices available to your account in the console.

Field reference

Online model catalog

Models within each vendor are ordered from newer to older versions. Use your browser find command or the search box above. Wide tables scroll horizontally on small screens.

OpenAI 92

Anthropic 10

Google 19

xAI 9

DeepSeek 8

阿里巴巴 77

字节跳动 10

智谱 10

Moonshot 5

MiniMax 2

Black Forest Labs 5

美团 1

阶跃星辰 1

Tiered pricing

Models marked with † use different price tiers based on the token count of a single request. The table below shows the input and output price for each tier. Confirm the final token count and charge in call logs.

Frequently asked questions

Is this price read in real time?

This page is generated periodically from current pricing configuration and shows its update time. Model changes, account groups, or dedicated contracts can produce a different console price; confirm final charges in the console and call logs.

When does the cache-read price apply?

Only when the request hits a supported prompt cache. If the model has no cache setting, the cache is missed, or the calling method does not support caching, normal input pricing applies.

How does tiered pricing work?

The system selects a tier from the token count of each request. The tier table lists both input and output prices; confirm final usage and charges in call logs.

How do I confirm enterprise purchasing, volume usage, or discounts?

Contact the site owner or support with the account, models, expected usage, contract, and invoice requirements. The documentation does not promise a fixed discount threshold or percentage.

How do I call a model after finding its ID?

Confirm the endpoint and token group shown in the table, then create the matching token in the console. OpenAI-compatible models generally use /v1/chat/completions or /v1/responses; Gemini, image, and specialized models should follow their dedicated documentation.

Pricing snapshot generated: 2026-09-02T04:12:17.377Z