Skip to main content

Model Catalog and Pricing

This page lists 281 online LaoZhang API models with input, output, cache-read, per-call, and tiered prices, token groups, and supported endpoints. The data was updated on 2026-07-25, 22:52 (UTC+8). Confirm final availability and charges in the signed-in console and call logs.

Online models281
Vendors13
Usage-based259
Per-call22
Confirming prices and next steps

Online models, cache-read prices, and tiered pricing rules are refreshed from current system configuration. Use the signed-in console to confirm account groups and current prices, and call logs to confirm actual charges.

Pricing notes

All amounts on this page are shown in US dollars. Usage-based input, output, and cache-read prices are per million tokens ($/1M). Per-call models are priced for each call. The table shows default list prices; account contracts or dedicated routes may use different prices.

Top-up bonus

A single top-up of US$700 or more receives a 5% balance bonus. The bonus does not rewrite the listed model price. Confirm the credited balance and actual charges in the console.

A model can support multiple token groups with different routes, permissions, or billing modes. Confirm the groups and actual prices available to your account in the console.

Field reference

Online model catalog

Models within each vendor are ordered from newer to older versions. Use your browser find command or the search box above. Wide tables scroll horizontally on small screens.

OpenAI 93

Anthropic 20

Google 20

xAI 30

DeepSeek 7

阿里巴巴 77

字节跳动 10

智谱 10

Moonshot 5

MiniMax 2

Black Forest Labs 5

美团 1

阶跃星辰 1

Tiered pricing

Models marked with † use different price tiers based on the token count of a single request. The table below shows the input and output price for each tier. Confirm the final token count and charge in call logs.

Frequently asked questions

Is this price read in real time?

This page is generated periodically from current pricing configuration and shows its update time. Model changes, account groups, or dedicated contracts can produce a different console price; confirm final charges in the console and call logs.

When does the cache-read price apply?

Only when the request hits a supported prompt cache. If the model has no cache setting, the cache is missed, or the calling method does not support caching, normal input pricing applies.

How does tiered pricing work?

The system selects a tier from the token count of each request. The tier table lists both input and output prices; confirm final usage and charges in call logs.

Does a US$700 top-up change the listed model price?

No. A qualifying single top-up receives a 5% balance bonus while the default model list price remains unchanged. Confirm the credited balance and charges in the console.

How do I call a model after finding its ID?

Confirm the endpoint and token group shown in the table, then create the matching token in the console. OpenAI-compatible models generally use /v1/chat/completions or /v1/responses; Gemini, image, and specialized models should follow their dedicated documentation.

Pricing snapshot generated: 2026-07-25T14:52:31.341Z