All models, one price list.

Transparent usage-based pricing across chat, embeddings, and image generation. Pay only for what you use.

One go-to model per family.

ModelTypePriceStatus
deepseek-r1
DeepSeek R1
Chat
¥3.72 in · ¥14.87 out
per 1M tokens
cache hit $0.7452 in
Available
kimi-k2.5
Kimi K2.5
Chat
¥3.72 in · ¥19.51 out
per 1M tokens
cache hit $0.7452 in
AvailableRetiring
Retiring on 2026-11-04. Switch to kimi-k2.5-cn.
qwen3.8-max
Qwen3.8 Max
Chat
¥10.69 in · ¥32.08 out
per 1M tokens
cache hit $1.3349 in
Available
qwen3-vl-8b-instruct-cn
Qwen3-VL 8B (北京)
Chat
¥0.47 in · ¥1.86 out
per 1M tokens
Available
qwen3-coder-30b-a3b-instruct-cn
Qwen3-Coder 30B (北京)
Chat
¥1.40 in · ¥5.58 out
per 1M tokens
Available
glm-5.2
GLM-5.2
Chat
¥7.13 in · ¥24.95 out
per 1M tokens
cache hit $1.7820 in
Available
minimax-m2.5
MiniMax M2.5
Chat
¥1.97 in · ¥7.86 out
per 1M tokens
cache hit $0.3953 in
AvailableRetiring
Retiring on 2026-11-04. Switch to minimax-m2.5-cn.

How billing works

Chat / embeddings bill per token; images bill per generation. Cost is debited at request time and recorded in your transactions, priced in CNY (¥).

Pay-as-you-go

No upfront commitment, no subscription. Top up any amount and use it across all models.

Flexible top-ups

Online top-ups (Alipay / WeChat Pay) are coming soon. In the meantime, contact support or your business contact to top up via corporate transfer.

Need higher rate limits or custom pricing for production scale? Email sales@tokengp.com — we'll work with you.

Start in 5 minutes.

Sign up, top up, ship. No card required.