All models, one price list.
Transparent usage-based pricing across chat, embeddings, and image generation. Pay only for what you use.
One go-to model per family.
| Model | Type | Price | Status |
|---|---|---|---|
deepseek-r1 DeepSeek R1 | Chat | ¥3.72 in · ¥14.87 out per 1M tokens cache hit $0.7452 in | Available |
kimi-k2.5 Kimi K2.5 | Chat | ¥3.72 in · ¥19.51 out per 1M tokens cache hit $0.7452 in | AvailableRetiring Retiring on 2026-11-04. Switch to kimi-k2.5-cn. |
qwen3.8-max Qwen3.8 Max | Chat | ¥10.69 in · ¥32.08 out per 1M tokens cache hit $1.3349 in | Available |
qwen3-vl-8b-instruct-cn Qwen3-VL 8B (北京) | Chat | ¥0.47 in · ¥1.86 out per 1M tokens | Available |
qwen3-coder-30b-a3b-instruct-cn Qwen3-Coder 30B (北京) | Chat | ¥1.40 in · ¥5.58 out per 1M tokens | Available |
glm-5.2 GLM-5.2 | Chat | ¥7.13 in · ¥24.95 out per 1M tokens cache hit $1.7820 in | Available |
minimax-m2.5 MiniMax M2.5 | Chat | ¥1.97 in · ¥7.86 out per 1M tokens cache hit $0.3953 in | AvailableRetiring Retiring on 2026-11-04. Switch to minimax-m2.5-cn. |
How billing works
Chat / embeddings bill per token; images bill per generation. Cost is debited at request time and recorded in your transactions, priced in CNY (¥).
Pay-as-you-go
No upfront commitment, no subscription. Top up any amount and use it across all models.
Flexible top-ups
Online top-ups (Alipay / WeChat Pay) are coming soon. In the meantime, contact support or your business contact to top up via corporate transfer.
Need higher rate limits or custom pricing for production scale? Email sales@tokengp.com — we'll work with you.