Supported Model Providers
Switch models per task — chat, reasoning, code, vision — without changing your client config.
General-purpose chat, content generation, translation, customer service.
DeepSeek · Qwen · Doubao · GPT · ClaudeDeep reasoning, document analysis, 1M-token context windows for book-length processing.
Kimi · GLM · StepFun · Claude OpusCode generation, tool calling, autonomous agent workflows, multi-step reasoning.
Qwen Coder · LongCat · MiMo · GPT CodexImage understanding, diagram analysis, OCR, image generation workflows.
Hunyuan · MiniMax · Doubao · GPT ImageHere's a realistic monthly task — so you know exactly what to expect.
Monthly volume
Input
2M
tokens
Output
1M
tokens
Actual cost varies by model multiplier and input/output ratio. Check the pricing table for real-time rates per model.
Three standard-compatible interfaces. Use whatever your client or SDK expects.
/v1/chat/completionsOpenAI-compatible. Drop-in replacement for any OpenAI SDK — just swap base URL.
/v1/responsesNext-gen response format. Ideal for agent workflows and structured tool calls.
/v1/messagesAnthropic Messages-compatible. Use with Claude Code, Cursor, or any Messages client.
Not sure which one? Use /v1/chat/completions — it works with the widest range of clients.
From DeepSeek to Claude, GPT to Grok — switch models instantly without changing keys. Pay only for what you use.
| Model | Provider | Price / 1M tokens | Context | Vision |
|---|---|---|---|---|
| claude-fable-5 | Anthropic | $2.39/M | 1M | ✅ |
| kimi-k3 | Moonshot / Kimi | $1.91/M | 1M | ✅ |
| claude-opus-5 | Anthropic | $1.53/M | 1M | ✅ |
| claude-opus-4-8 | Anthropic | $1.53/M | 1M | ✅ |
| claude-opus-4-7 | Anthropic | $1.53/M | 1M | ✅ |
| claude-opus-4-6 | Anthropic | $1.53/M | 1M | ✅ |
| claude-sonnet-5 | Anthropic | $0.92/M | 1M | ✅ |
| qwen3.8-max | Alibaba / Qwen | $0.77/M | 1M | ✅ |
| glm-5.2 | Zhipu / Z.ai | $0.61/M | 1M | ✅ |
| qwen3.7-max | Alibaba / Qwen | $0.61/M | 1M | — |
| gpt-5.6-sol | OpenAI | $0.61/M | 258K | ✅ |
| gpt-5.5 | OpenAI | $0.61/M | 258K | ✅ |
| gpt-5.6-luna | OpenAI | $0.46/M | 258K | ✅ |
| gpt-5.6-terra | OpenAI | $0.46/M | 258K | ✅ |
| deepseek-v4-pro | DeepSeek | $0.44/M | 1M | — |
| glm-5.1 | Zhipu / Z.ai | $0.38/M | 256K | ✅ |
| qwen3.7-plus | Alibaba / Qwen | $0.31/M | 1M | ✅ |
| gpt-5.4 | OpenAI | $0.31/M | 1M | ✅ |
| qwen3.6-plus | Alibaba / Qwen | $0.23/M | 1M | ✅ |
| doubao-seed-2.0-code | ByteDance / Doubao | $0.23/M | 200K | ✅ |
| doubao-seed-2.0-pro | ByteDance / Doubao | $0.23/M | 128K | ✅ |
| mimo-v2.5-pro | Xiaomi / MiMo | $0.23/M | 1M | — |
| grok-4.5 | xAI / Grok | $0.23/M | 500K | ✅ |
| mimo-v2.5 | Xiaomi / MiMo | $0.15/M | 1M | ✅ |
| deepseek-v4-flash | DeepSeek | $0.15/M | 1M | — |
| claude-haiku-4-5-20251001 | Anthropic | $0.08/M | 256K | ✅ |
| LongCat-2.0 | Meituan / LongCat | $0.08/M | 1M | — |
| hy3 | Tencent Hunyuan | $0.08/M | 256K | — |
| MiniMax-M3 | MiniMax | $0.08/M | 1M | ✅ |
| step-3.7-flash | StepFun | $0.08/M | 256K | ✅ |
| gpt-5.3-codex-spark | OpenAI / Codex | $0.08/M | 128K | — |
| claude-sonnet-4-6HOT | AI API Proxy | $0.05/M | 1M | ✅ |
Anonymized from real support conversations.
“Internal knowledge base connected seamlessly. The team only maintains one API config, and everyone keeps using their preferred clients.”
“Customer QA, product copy, and translation all share one Key. Model switching per task makes monthly cost tracking straightforward.”
“Long-doc extraction, code assist, and reasoning tasks need different models. Shared balance saved students from applying for multiple accounts.”
“Delivering AI workflows to clients with both OpenAI-compatible and Messages endpoints. Switching clients doesn't require redoing the entire config.”
Not just an API. Everything you need to use it in production.
We do not log request content. Your data is forwarded, never stored.
Contact support with your order number for official invoicing.
WhatsApp instant support. Topping-up issues, refunds — we handle it.
Common questions, answered upfront.
Register an account, top up credits, and create an API Key in your dashboard. Keys are instantly active.
Use https://your-domain.com/api/v1 as your base URL. Append /chat/completions, /responses, or /messages depending on your protocol.
No. Credits never expire. Top up what you need, when you need it. No subscription, no minimum.
30+ models including DeepSeek, Qwen, Kimi, GLM, GPT, Claude, Grok, Doubao, Hunyuan, MiniMax, LongCat, StepFun, and MiMo. Full list on the pricing page.
Each model has a base price per 1M tokens and a multiplier. Your cost = basePrice × multiplier × tokens. Check the pricing table for real-time rates.
Check that your API Key is correct (no extra spaces), the Authorization header uses 'Bearer', and the base URL is correct. Contact support if it persists.
More questions? View Full FAQ or Contact Support