AIWave field guide / Sep 12, 2026

Chinese AI Model Pricing Tracker (September 2026)

A dated, filter-ready table of AIWave base rates for DeepSeek, GLM, Kimi, ERNIE, MiniMax, Qwen, Doubao, StepFun, and MiMo routes.

OpenAI-compatible routeDated pricingRunnable code

Snapshot generated on Sep 12, 2026 from AIWave's public pricing JSON. The snapshot's latest checked date is 2026-09-10. Each row keeps its own effective date.

This tracker puts input, cached input, and output on separate lines of thought. That matters because an output-heavy coding agent and an input-heavy document pipeline can reverse the ranking you get from input price alone.

How to read the table

Rates are per 1M text tokensInput and output use the same unit, which makes a workload calculation explicit.
Cache is a separate field"Not listed" means the public rate card does not publish a cache-hit rate for that route.
Dates belong to rowsA new row can have a newer effective date than the rest of the catalog.
Account group still appliesThese are dated base rates before the effective account-group multiplier.

The provider families in this snapshot are DeepSeek, Doubao, ERNIE, GLM, Kimi, MiMo, MiniMax, Qwen, StepFun. The table is catalog data, not a claim that every capability or route has been tested for every workload.

Dated base rates

ProviderModel IDInput / 1MCached input / 1MOutput / 1MEffective
DeepSeekdeepseek-flash$0.7$0.0233$2.12026-09-10
DeepSeekdeepseek-v3.2$0.154Not listed$0.3082026-08-27
DeepSeekdeepseek-v4-flash$0.638$0.020288$1.9142026-08-27
DeepSeekdeepseek-v4-pro$1.914$0.063736$5.7422026-08-27
Doubaodoubao-seed-2-0-lite-260428$0.401708Not listed$2.4102482026-08-27
Doubaodoubao-seed-2-0-mini-260428$0.178537Not listed$1.7853692026-08-27
Doubaodoubao-seed-2-1-pro-260628$1.339027Not listed$6.6951332026-08-27
Doubaodoubao-seed-evolving$1.339027Not listed$6.6951332026-08-27
ERNIEernie-4.5-turbo$0.19726$0.049315$0.7452092026-08-27
ERNIEernie-5.0$2.465754$0.246575$9.3151252026-08-27
GLMglm-4.5$0.6975$0.18$2.172026-08-27
GLMglm-4.5-air$0.6525$0.162$2.032026-08-27
GLMglm-4.6$0.93$0.22$3.412026-08-27
GLMglm-4.7$0.93$0.22$3.412026-08-27
GLMglm-5$1.55$0.400001$4.962026-08-27
GLMglm-5-turbo$1.8$0.480001$5.42026-08-27
GLMglm-5.1$2.1$0.680001$6.62026-08-27
Kimikimi-k2.5$0.66$0.121846$3.32026-08-27
Kimikimi-k2.6$1.09$0.186857$4.59982026-08-27
Kimikimi-k2.7-code$1.89$0.285001$62026-08-27
Kimikimi-k2.7-code-highspeed$4.725$0.712502$14.9999992026-08-27
Kimikimi-k3$4.5$0.9$22.52026-08-27
Kimimoonshot-v1-128k$1.8Not listed$4.52026-08-27
Kimimoonshot-v1-128k-vision-preview$2.3Not listed$5.752026-08-27
Kimimoonshot-v1-32k$0.95Not listed$2.852026-08-27
Kimimoonshot-v1-32k-vision-preview$1.15Not listed$3.452026-08-27
Kimimoonshot-v1-8k$0.3Not listed$2.22026-08-27
Kimimoonshot-v1-8k-vision-preview$0.3Not listed$2.3012026-08-27
Kimimoonshot-v1-auto$75Not listed$752026-08-27
MiMoxiaomi/mimo-v2.5-pro$1.562198Not listed$4.6865932026-08-27
MiniMaxMiniMax-M2$0.46866$0.046866$1.874642026-08-27
MiniMaxMiniMax-M2.1$0.46866$0.046866$1.874642026-08-27
MiniMaxMiniMax-M2.1-highspeed$0.93732$0.046866$3.749282026-08-27
MiniMaxMiniMax-M2.5$0.46866$0.046866$1.874642026-08-27
MiniMaxMiniMax-M2.5-highspeed$0.93732$0.046866$3.749282026-08-27
MiniMaxMiniMax-M2.7$0.45304$0.090608$1.812162026-08-27
MiniMaxMiniMax-M2.7-highspeed$0.93732$0.093732$3.749282026-08-27
MiniMaxMiniMax-M3$0.90608$0.181216$3.624322026-08-27
Qwenqwen-72b-chat$4.463422Not listed$4.4634222026-08-27
Qwenqwen-audio-3.0-realtime-flash$0.669513Not listed$0.6695132026-08-27
Qwenqwen-audio-3.0-realtime-plus$1.115856Not listed$1.1158562026-08-27
Qwenqwen-image-2.0-pro-2026-04-22$0.122744Not listed$0.1227442026-08-27
Qwenqwen-image-3.0$0.040171Not listed$0.0401712026-08-27
Qwenqwen-image-3.0-pro$0.055793Not listed$0.0557932026-08-27
Qwenqwen1.5-72b-chat$4.463422Not listed$4.4634222026-08-27
Qwenqwen3-max$1.562198Not listed$6.2487912026-08-27
Qwenqwen3.5-122b-a10b$0.49315Not listed$3.7260442026-08-27
Qwenqwen3.5-27b$0.443836Not listed$3.3534472026-08-27
Qwenqwen3.5-omni-flash$0.490976Not listed$2.9681762026-08-27
Qwenqwen3.5-plus$0.446342Not listed$2.6780532026-08-27
Qwenqwen3.6-27b$0.669513Not listed$4.017082026-08-27
Qwenqwen3.6-35b-a3b$0.401708Not listed$2.4102482026-08-27
Qwenqwen3.6-flash$0.267805Not listed$1.6068322026-08-27
Qwenqwen3.7-flash-2026-07-15$0.267805Not listed$1.0712212026-08-27
Qwenqwen3.7-max-2026-05-20$2.678053Not listed$8.034162026-08-27
Qwenqwen3.7-max-2026-06-08$2.678053Not listed$8.034162026-08-27
Qwenqwen3.7-plus$0.446342Not listed$1.7853692026-08-27
Qwenqwen3.7-text-embedding$0.111586Not listed$0.1115862026-08-27
Qwenqwen3.8-2.4t-a95b$2.678053Not listed$8.034162026-08-27
Qwenqwen3.8-27b$0.669513Not listed$2.6780532026-08-27
Qwenqwen3.8-max$2.678053Not listed$8.034162026-08-27
StepFunstep-3.5-flash$0.21$0.042$0.632026-08-27
StepFunstep-3.5-flash-2603$0.21$0.042$0.632026-08-27
StepFunstep-3.7-flash$0.4$0.08$2.42026-08-27

Forecast a workload

Start with a token mix, not a model label. For 50,000 uncached input tokens and 2,000 output tokens, multiply each token count by the matching per-million rate. Add cached input only when the ledger reports a cache hit.

input_cost = input_tokens / 1_000_000 * input_rate
cache_cost = cached_input_tokens / 1_000_000 * cache_rate
output_cost = output_tokens / 1_000_000 * output_rate
base_cost = input_cost + cache_cost + output_cost

Run one controlled request after choosing a route. Compare the model ID, token fields, account group, and charged quota in the request ledger with your estimate.

What the table does not prove

Price does not establish quality, context capacity, tool-call support, or route availability. Check the model catalog and current status page, then run a representative test set. Some capability fields in the internal catalog are intentionally marked unknown until they have direct evidence.

Use the JSON endpoint for automation. Use this page when a person needs to scan the catalog, compare token categories, and retain the effective date beside a forecast.

Related AIWave links