Pricing and Billing
What you are charged for, how caching lowers your cost, and how to check any single call against your bill.
1. What you pay for
Pricing is per token, quoted per 1 million tokens, and charged in CNY (¥). There is no subscription and no monthly minimum — you buy credits, and each call deducts what it actually used. Credits never expire.
Every call is billed across up to four separate token types, each with its own rate. This matters more than it sounds: output is typically 5–6× the input rate on the same model, so a rough estimate based on input alone will understate your bill.
| Token type | Typical rate | What it is |
|---|---|---|
| Input | Baseline | The prompt you send, minus anything served from cache. |
| Output | 5–6× input on most models | What the model generates. Usually the largest line on your bill — cap it with max_tokens if cost matters more than length. A few models price output far lower, so check the model you actually use. |
| Cache read | ~10% of input on most models | Prompt content the provider already had cached from an earlier call. Heavily discounted — but not on every model: on a few, cached input costs the same as or more than regular input, so caching is not automatically a saving. |
| Cache write | Varies by model | Content being placed into the cache so later calls can reuse it. On Claude models this runs above the input rate; on several others it is below. Per-model rates are on the pricing page. |
One API key reaches every model in your key's scope — you do not need a separate key per model. All of your keys draw on the same account balance.
2. All model prices
Every rate below is what your account is actually charged — these come from the same price records the billing system reads, so nothing here can drift out of sync with your bill. Prices are shown in USD, converted from CNY at 6.72 CNY/USD.
Models available
53
Providers
14
Lowest input price
$0.0387 / 1M
Exchange rate
6.72 CNY/USD
Showing 53 of 53 models
| Model | Provider | Input / 1M | Output / 1M | Cache read / 1M | Cache write / 1M | Context | Images | Copy model ID |
|---|---|---|---|---|---|---|---|---|
| JZS Token1 model | ||||||||
| jzs-max-3.0Exclusive | JZS Token | $0.0619 | $0.3095 | $0.0062 | $0.0619 | 1M | Understands | |
| Alibaba / Qwen7 models | ||||||||
| qwen3.6-flash | Alibaba / Qwen | $0.154838% below official | $0.7738 | $0.0155 | $0.1548 | — | Unknown | |
| qwen3.8-flash | Alibaba / Qwen | $0.1548 | $0.7738 | $0.0155 | $0.1548 | — | Unknown | |
| qwen3.6-plus | Alibaba / Qwen | $0.232153% below official | $1.16 | $0.0232 | $0.2321 | 1M | Understands | |
| qwen3.7-plus | Alibaba / Qwen | $0.3095 | $1.55 | $0.0310 | $0.3095 | 1M | Understands | |
| qwen3.8-omni-flash | Alibaba / Qwen | $0.3869 | $1.93 | $0.0387 | $0.3869 | — | Unknown | |
| qwen3.7-max | Alibaba / Qwen | $0.773869% below official | $2.32 | $0.0967 | $0.9673 | 1M | — | |
| qwen3.8-max | Alibaba / Qwen | $1.5522% below official | $4.64 | $0.1935 | $1.93 | 1M | Understands | |
| Anthropic7 models | ||||||||
| claude-haiku-4-5-20251001 | Anthropic | $0.077492% below official | $0.3869 | $0.0077 | $0.1161 | 256K | Understands | |
| claude-sonnet-5 | Anthropic | $0.773861% below official | $3.87 | $0.0774 | $0.9673 | 1M | Understands | |
| claude-opus-4-6 | Anthropic | $1.16 | $5.80 | $0.1161 | $1.45 | 1M | Understands | |
| claude-opus-4-7 | Anthropic | $1.16 | $5.80 | $0.1161 | $1.45 | 1M | Understands | |
| claude-opus-4-8 | Anthropic | $1.16 | $5.80 | $0.1161 | $1.45 | 1M | Understands | |
| claude-opus-5 | Anthropic | $1.1676% below official | $5.80 | $0.1161 | $1.45 | 1M | Understands | |
| claude-opus-4-5-20251101 | Anthropic | $1.5569% below official | $7.74 | $0.0155 | $1.93 | — | Unknown | |
| ByteDance / Doubao4 models | ||||||||
| doubao-seed-2.0-lite | ByteDance / Doubao | $0.07748% below official | $0.3869 | $0.0077 | $0.0387 | — | Unknown | |
| doubao-seed-2.0-code | ByteDance / Doubao | $0.232148% below official | $1.16 | $0.0232 | $0.2321 | 200K | Understands | |
| doubao-seed-2.0-pro | ByteDance / Doubao | $0.232148% below official | $1.16 | $0.0232 | $0.2321 | 128K | Understands | |
| doubao-seed-2.1-turbo | ByteDance / Doubao | $0.232145% below official | $1.16 | $0.0232 | $0.2321 | — | Unknown | |
| DeepSeek3 models | ||||||||
| deepseek-v4-flash⏱ Off-peak rate · peak 09:00–12:00 ×2, 14:00–18:00 ×2 (Asia/Shanghai) | DeepSeek | $0.1488 | $0.7440 | $0.0149 | $0.1488 | 1M | — | |
| deepseek-v4.1-flash⏱ Off-peak rate · peak 09:00–12:00 ×2, 14:00–18:00 ×2 (Asia/Shanghai) | DeepSeek | $0.1488 | $0.7440 | $0.0149 | $0.1488 | — | Unknown | |
| deepseek-v4-pro⏱ Off-peak rate · peak 09:00–12:00 ×2, 14:00–18:00 ×2 (Asia/Shanghai) | DeepSeek | $0.4464 | $2.23 | $0.0446 | $0.4464 | 1M | — | |
| Google / Gemini3 models | ||||||||
| gemini-3-flash | Google / Gemini | $0.2321 | $1.16 | $0.0232 | $0.1161 | — | Unknown | |
| gemini-3.5-flash | Google / Gemini | $0.4643 | $2.32 | $0.0464 | $0.2321 | — | Unknown | |
| gemini-3.1-pro | Google / Gemini | $0.7738 | $3.87 | $0.0774 | $0.3869 | — | Unknown | |
| Meituan / LongCat1 model | ||||||||
| LongCat-2.0 | Meituan / LongCat | $0.077474% below official | $0.3095 | $0.0077 | $0.0774 | 1M | — | |
| MiniMax4 models | ||||||||
| MiniMax-M2.7 | MiniMax | $0.038783% below official | $0.1935 | $0.0039 | $0.0193 | — | Unknown | |
| MiniMax-M2.7-highspeed | MiniMax | $0.077483% below official | $0.3869 | $0.0077 | $0.0387 | — | Unknown | |
| MiniMax-M3 | MiniMax | $0.077467% below official | $0.3869 | $0.0077 | $0.0193 | 1M | Understands | |
| MiniMax-M3-highspeed | MiniMax | $0.1548 | $0.7738 | $0.0155 | $0.0774 | — | Unknown | |
| Moonshot / Kimi3 models | ||||||||
| kimi-k2.6 | Moonshot / Kimi | $0.154880% below official | $0.7738 | $0.0155 | $0.1548 | — | Unknown | |
| kimi-k2.7 | Moonshot / Kimi | $0.4643 | $2.32 | $0.0464 | $0.4643 | — | Unknown | |
| kimi-k3 | Moonshot / Kimi | $1.9335% below official | $9.67 | $0.1935 | $1.93 | 1M | Understands | |
| OpenAI9 models | ||||||||
| gpt-6-luna | OpenAI | $0.0387 | $0.2321 | $0.0039 | $0.0484 | — | Unknown | |
| gpt-5.6-luna | OpenAI | $0.0774 | $0.0464 | $0.0077 | $0.1161 | 258K | Understands | |
| gpt-5.6-terra | OpenAI | $0.2321 | $1.39 | $0.0232 | $0.3482 | 258K | Understands | |
| gpt-6-sol | OpenAI | $0.2321 | $1.39 | $0.0232 | $0.2902 | — | Unknown | |
| gpt-5.5 | OpenAI | $0.3869 | $2.32 | $0.0387 | $0.3869 | 258K | Understands | |
| gpt-5.6-sol | OpenAI | $0.4643 | $2.79 | $0.0464 | $0.5804 | 258K | Understands | |
| gpt-6-astra | OpenAI | $1.16 | $6.96 | $0.1161 | $1.74 | — | Unknown | |
| gpt-image-2-4k | OpenAI | $0.1935 per image | 4K output | Generates | ||||
| gpt-image-2 | OpenAI | $0.0193 per image | 1–2K output | Generates | ||||
| StepFun1 model | ||||||||
| step-3.7-flash | StepFun | $0.077459% below official | $0.3869 | $0.0077 | $0.0387 | 256K | Understands | |
| xAI / Grok3 models | ||||||||
| grok-4.5 | xAI / Grok | $0.3869 | $1.93 | $0.4836 | $0.3869 | 500K | Understands | |
| grok-4.6 | xAI / Grok | $0.3869 | $1.93 | $0.4836 | $0.3869 | — | Unknown | |
| grok-4.7 | xAI / Grok | $0.3869 | $1.93 | $0.0387 | $0.4836 | — | Unknown | |
| Xiaomi / MiMo4 models | ||||||||
| mimo-v2.5 | Xiaomi / MiMo | $0.1548 | $0.7738 | $0.0155 | $0.1548 | 1M | Understands | |
| mimo-v2.6-flash | Xiaomi / MiMo | $0.1548 | $0.7738 | $0.0155 | $0.1548 | — | Unknown | |
| mimo-v2.5-pro | Xiaomi / MiMo | $0.3095 | $1.55 | $0.0310 | $0.3095 | 1M | — | |
| mimo-v2.6-pro | Xiaomi / MiMo | $0.3095 | $1.55 | $0.0310 | $0.3095 | — | Unknown | |
| Zhipu / Z.ai3 models | ||||||||
| glm-5.3-flash | Zhipu / Z.ai | $0.1548 | $0.7738 | $0.0155 | $0.1548 | — | Unknown | |
| glm-5.2 | Zhipu / Z.ai | $0.619029% below official | $3.10 | $0.0619 | $0.6190 | 1M | Understands | |
| glm-5.3 | Zhipu / Z.ai | $0.619029% below official | $3.10 | $0.0619 | $0.6190 | — | Unknown | |
JZS Token1 model
JZS Token
- Input / 1M
- $0.0619
- Output / 1M
- $0.3095
- Cache read
- $0.0062
- Cache write
- $0.0619
Alibaba / Qwen7 models
Alibaba / Qwen
- Input / 1M
- $0.1548
- Output / 1M
- $0.7738
- Cache read
- $0.0155
- Cache write
- $0.1548
38% below the vendor's list price
Alibaba / Qwen
- Input / 1M
- $0.1548
- Output / 1M
- $0.7738
- Cache read
- $0.0155
- Cache write
- $0.1548
Alibaba / Qwen
- Input / 1M
- $0.2321
- Output / 1M
- $1.16
- Cache read
- $0.0232
- Cache write
- $0.2321
53% below the vendor's list price
Alibaba / Qwen
- Input / 1M
- $0.3095
- Output / 1M
- $1.55
- Cache read
- $0.0310
- Cache write
- $0.3095
Alibaba / Qwen
- Input / 1M
- $0.3869
- Output / 1M
- $1.93
- Cache read
- $0.0387
- Cache write
- $0.3869
Alibaba / Qwen
- Input / 1M
- $0.7738
- Output / 1M
- $2.32
- Cache read
- $0.0967
- Cache write
- $0.9673
69% below the vendor's list price
Alibaba / Qwen
- Input / 1M
- $1.55
- Output / 1M
- $4.64
- Cache read
- $0.1935
- Cache write
- $1.93
22% below the vendor's list price
Anthropic7 models
Anthropic
- Input / 1M
- $0.0774
- Output / 1M
- $0.3869
- Cache read
- $0.0077
- Cache write
- $0.1161
92% below the vendor's list price
Anthropic
- Input / 1M
- $0.7738
- Output / 1M
- $3.87
- Cache read
- $0.0774
- Cache write
- $0.9673
61% below the vendor's list price
Anthropic
- Input / 1M
- $1.16
- Output / 1M
- $5.80
- Cache read
- $0.1161
- Cache write
- $1.45
Anthropic
- Input / 1M
- $1.16
- Output / 1M
- $5.80
- Cache read
- $0.1161
- Cache write
- $1.45
Anthropic
- Input / 1M
- $1.16
- Output / 1M
- $5.80
- Cache read
- $0.1161
- Cache write
- $1.45
Anthropic
- Input / 1M
- $1.16
- Output / 1M
- $5.80
- Cache read
- $0.1161
- Cache write
- $1.45
76% below the vendor's list price
Anthropic
- Input / 1M
- $1.55
- Output / 1M
- $7.74
- Cache read
- $0.0155
- Cache write
- $1.93
69% below the vendor's list price
ByteDance / Doubao4 models
ByteDance / Doubao
- Input / 1M
- $0.0774
- Output / 1M
- $0.3869
- Cache read
- $0.0077
- Cache write
- $0.0387
8% below the vendor's list price
ByteDance / Doubao
- Input / 1M
- $0.2321
- Output / 1M
- $1.16
- Cache read
- $0.0232
- Cache write
- $0.2321
48% below the vendor's list price
ByteDance / Doubao
- Input / 1M
- $0.2321
- Output / 1M
- $1.16
- Cache read
- $0.0232
- Cache write
- $0.2321
48% below the vendor's list price
ByteDance / Doubao
- Input / 1M
- $0.2321
- Output / 1M
- $1.16
- Cache read
- $0.0232
- Cache write
- $0.2321
45% below the vendor's list price
DeepSeek3 models
DeepSeek
- Input / 1M
- $0.1488
- Output / 1M
- $0.7440
- Cache read
- $0.0149
- Cache write
- $0.1488
DeepSeek
- Input / 1M
- $0.1488
- Output / 1M
- $0.7440
- Cache read
- $0.0149
- Cache write
- $0.1488
DeepSeek
- Input / 1M
- $0.4464
- Output / 1M
- $2.23
- Cache read
- $0.0446
- Cache write
- $0.4464
Google / Gemini3 models
Google / Gemini
- Input / 1M
- $0.2321
- Output / 1M
- $1.16
- Cache read
- $0.0232
- Cache write
- $0.1161
Google / Gemini
- Input / 1M
- $0.4643
- Output / 1M
- $2.32
- Cache read
- $0.0464
- Cache write
- $0.2321
Google / Gemini
- Input / 1M
- $0.7738
- Output / 1M
- $3.87
- Cache read
- $0.0774
- Cache write
- $0.3869
Meituan / LongCat1 model
Meituan / LongCat
- Input / 1M
- $0.0774
- Output / 1M
- $0.3095
- Cache read
- $0.0077
- Cache write
- $0.0774
74% below the vendor's list price
MiniMax4 models
MiniMax
- Input / 1M
- $0.0387
- Output / 1M
- $0.1935
- Cache read
- $0.0039
- Cache write
- $0.0193
83% below the vendor's list price
MiniMax
- Input / 1M
- $0.0774
- Output / 1M
- $0.3869
- Cache read
- $0.0077
- Cache write
- $0.0387
83% below the vendor's list price
MiniMax
- Input / 1M
- $0.0774
- Output / 1M
- $0.3869
- Cache read
- $0.0077
- Cache write
- $0.0193
67% below the vendor's list price
MiniMax
- Input / 1M
- $0.1548
- Output / 1M
- $0.7738
- Cache read
- $0.0155
- Cache write
- $0.0774
Moonshot / Kimi3 models
Moonshot / Kimi
- Input / 1M
- $0.1548
- Output / 1M
- $0.7738
- Cache read
- $0.0155
- Cache write
- $0.1548
80% below the vendor's list price
Moonshot / Kimi
- Input / 1M
- $0.4643
- Output / 1M
- $2.32
- Cache read
- $0.0464
- Cache write
- $0.4643
Moonshot / Kimi
- Input / 1M
- $1.93
- Output / 1M
- $9.67
- Cache read
- $0.1935
- Cache write
- $1.93
35% below the vendor's list price
OpenAI9 models
OpenAI
- Input / 1M
- $0.0387
- Output / 1M
- $0.2321
- Cache read
- $0.0039
- Cache write
- $0.0484
OpenAI
- Input / 1M
- $0.0774
- Output / 1M
- $0.0464
- Cache read
- $0.0077
- Cache write
- $0.1161
OpenAI
- Input / 1M
- $0.2321
- Output / 1M
- $1.39
- Cache read
- $0.0232
- Cache write
- $0.3482
OpenAI
- Input / 1M
- $0.2321
- Output / 1M
- $1.39
- Cache read
- $0.0232
- Cache write
- $0.2902
OpenAI
- Input / 1M
- $0.3869
- Output / 1M
- $2.32
- Cache read
- $0.0387
- Cache write
- $0.3869
OpenAI
- Input / 1M
- $0.4643
- Output / 1M
- $2.79
- Cache read
- $0.0464
- Cache write
- $0.5804
OpenAI
- Input / 1M
- $1.16
- Output / 1M
- $6.96
- Cache read
- $0.1161
- Cache write
- $1.74
StepFun1 model
StepFun
- Input / 1M
- $0.0774
- Output / 1M
- $0.3869
- Cache read
- $0.0077
- Cache write
- $0.0387
59% below the vendor's list price
xAI / Grok3 models
xAI / Grok
- Input / 1M
- $0.3869
- Output / 1M
- $1.93
- Cache read
- $0.4836
- Cache write
- $0.3869
xAI / Grok
- Input / 1M
- $0.3869
- Output / 1M
- $1.93
- Cache read
- $0.4836
- Cache write
- $0.3869
xAI / Grok
- Input / 1M
- $0.3869
- Output / 1M
- $1.93
- Cache read
- $0.0387
- Cache write
- $0.4836
Xiaomi / MiMo4 models
Xiaomi / MiMo
- Input / 1M
- $0.1548
- Output / 1M
- $0.7738
- Cache read
- $0.0155
- Cache write
- $0.1548
Xiaomi / MiMo
- Input / 1M
- $0.1548
- Output / 1M
- $0.7738
- Cache read
- $0.0155
- Cache write
- $0.1548
Xiaomi / MiMo
- Input / 1M
- $0.3095
- Output / 1M
- $1.55
- Cache read
- $0.0310
- Cache write
- $0.3095
Xiaomi / MiMo
- Input / 1M
- $0.3095
- Output / 1M
- $1.55
- Cache read
- $0.0310
- Cache write
- $0.3095
Zhipu / Z.ai3 models
Zhipu / Z.ai
- Input / 1M
- $0.1548
- Output / 1M
- $0.7738
- Cache read
- $0.0155
- Cache write
- $0.1548
Zhipu / Z.ai
- Input / 1M
- $0.6190
- Output / 1M
- $3.10
- Cache read
- $0.0619
- Cache write
- $0.6190
29% below the vendor's list price
Zhipu / Z.ai
- Input / 1M
- $0.6190
- Output / 1M
- $3.10
- Cache read
- $0.0619
- Cache write
- $0.6190
29% below the vendor's list price
Last updated Sep 24, 2026, 10:20 AM UTC. The pricing page shows the same catalogue with model comparisons and copyable model IDs.
3. Caching, and how to actually benefit from it
When consecutive requests share a long identical prefix, the provider can reuse the work it already did on that prefix instead of reprocessing it. You are charged the discounted cache-read rate for that portion. Long conversations, follow-up questions on the same document, code completion, and agent loops all hit this case naturally.
The practical rule: keep the unchanging part of your prompt at the front, and put what varies at the end. Caching matches on a shared prefix, so a system prompt or document that sits before your question stays cacheable across turns. Move a timestamp or a request ID to the top and you invalidate the prefix on every single call — you then pay full input rate for the whole thing, plus a cache write.
Cache writes are not free, so a one-off call that will never be followed up gains nothing from caching. The benefit shows up from the second call onward, which is exactly the pattern agents and multi-turn chat produce.
4. Image models are billed per image
Image generation is not billed per token. Each image is a fixed price regardless of prompt length, and the rate depends on the output resolution. Current per-image prices are on the pricing page.
Image calls go to POST /v1/images/generations, not the chat endpoint. In your usage records they show a cost with zero tokens — that is expected, since tokens play no part in what you were charged.
5. Checking a call against your bill
The usage lookup page takes an API key and shows your balance, per-model totals, and recent calls. The four token types appear as separate columns, so you can multiply each one by its rate on the pricing page and land on the exact figure that was deducted.
The token counts come from the provider's usage block on the response — the same numbers your own client receives, so you can verify them against your logs without asking us.
Failed calls are listed too, with the reason in the Details column and a cost of zero. A call that errored is never billed. If the failure came from the upstream provider rather than your request, that column says so — you do not need to go debugging your own code first.
6. When prices change
Model prices track what the providers charge, and those move occasionally. Our price records sync automatically and the pricing page shows when they were last updated. Each call is billed at the rate in effect at that moment, and your usage records keep the resulting cost, so a later price change never restates what you already paid.
A few models are priced by time of day, meaning a call during the provider's peak window costs more than the same call off-peak. Where that applies, the pricing page notes it on the model. If your workload is flexible, shifting it out of the peak window is the simplest saving available.
Questions about a specific charge
Send us the call's timestamp and the model, plus the masked form of your API key — never a full working key. See the contact page for how to reach us.