All models, one price list.
Transparent usage-based pricing across chat, embeddings, and image generation. Pay only for what you use.
One go-to model per family.
| Model | Type | Price | Status |
|---|---|---|---|
deepseek-v4-pro DeepSeek v4 Pro | Chat | ¥10.80 in · ¥21.60 out per 1M tokens cache hit ¥0.9000 in | Available |
kimi-k2.5 Kimi K2.5 | Chat | ¥3.60 in · ¥18.90 out per 1M tokens cache hit ¥0.7200 in | Available |
qwen3.8-max Qwen3.8 Max | Chat | ¥10.80 in · ¥32.40 out per 1M tokens cache hit ¥1.3500 in | Available |
qwen3-vl-plus Qwen3-VL Plus | Chat | ¥0.90 in · ¥9.00 out+ per 1M tokens cache hit ¥0.1800 in | Available |
qwen3-coder-flash Qwen3 Coder Flash | Chat | ¥0.90 in · ¥3.60 out+ per 1M tokens cache hit ¥0.1800 in | Available |
glm-5.3 GLM-5.3 | Chat | ¥7.20 in · ¥25.20 out per 1M tokens cache hit ¥1.8000 in | Available |
minimax-m2.5 MiniMax M2.5 | Chat | ¥1.89 in · ¥7.56 out per 1M tokens cache hit ¥0.3780 in | Available |
How billing works
Chat / embeddings bill per token; images bill per generation. Cost is debited at request time and recorded in your transactions, priced in CNY (¥).
Pay-as-you-go
No upfront commitment, no subscription. Top up any amount and use it across all models.
Flexible top-ups
Online top-ups (Alipay / WeChat Pay) are coming soon. In the meantime, contact support or your business contact to top up via corporate transfer.
Need higher rate limits or custom pricing for production scale? Email sales@tokengp.com — we'll work with you.