LLM price barometer
Every list-price move measured between 26 Sept – 3 Oct, biggest first. Percentages are on the blended 30/70 input/output price. Data snapshot: 2026-10-03.
Got more expensive (13)
- +291.3% DeepSeek V4 Flash 0731 (DeepSeek) — blended list price $0.23 to $0.90 per 1M tokens; input -18.6%, output +300.0%. Large move: worth confirming against the vendor's own rate card before you act on it. Input and output moved in opposite directions, so the effect on your bill depends on your token mix.
- +269.0% GLM 5.3 (Zhipu) — blended list price $0.95 to $3.50 per 1M tokens; input +269.0%, output +269.0%. Not a repricing: this model is sold by several hosts at different rates, and the one we quote changed. The previous rate was still on sale elsewhere. Large move: worth confirming against the vendor's own rate card before you act on it.
- +150.0% DeepSeek V4 Pro 0813 (DeepSeek) — blended list price $0.63 to $1.58 per 1M tokens; input +150.0%, output +150.0%. Not a repricing: this model is sold by several hosts at different rates, and the one we quote changed. The previous rate was still on sale elsewhere. Large move: worth confirming against the vendor's own rate card before you act on it.
- +79.6% GLM 5.2 (Zhipu) — blended list price $1.62 to $2.92 per 1M tokens; input -36.9%, output +95.4%. Not a repricing: this model is sold by several hosts at different rates, and the one we quote changed. The previous rate was still on sale elsewhere. Input and output moved in opposite directions, so the effect on your bill depends on your token mix.
- +76.2% GLM 5.3 Flash (Zhipu) — blended list price $0.36 to $0.64 per 1M tokens; input -34.3%, output +80.0%. Input and output moved in opposite directions, so the effect on your bill depends on your token mix.
- +33.3% Gemma 4 26B A4B (Google) — blended list price $0.18 to $0.24 per 1M tokens; input +33.3%, output +33.3%.
- +25.0% Grok 4.7 (xAI) — blended list price $3.84 to $4.80 per 1M tokens; input +25.0%, output +25.0%.
- +24.2% Muse Glimmer 30B (meta) — blended list price $0.93 to $1.15 per 1M tokens; input +16.7%, output +25.0%. Not a repricing: this model is sold by several hosts at different rates, and the one we quote changed. The previous rate was still on sale elsewhere.
- +14.2% DeepSeek V3 0324 (DeepSeek) — blended list price $0.77 to $0.88 per 1M tokens; input +16.0%, output +14.0%. Not a repricing: this model is sold by several hosts at different rates, and the one we quote changed. The previous rate was still on sale elsewhere.
- +10.9% DeepSeek V3 (DeepSeek) — blended list price $0.72 to $0.80 per 1M tokens; input -19.6%, output +15.6%. Not a repricing: this model is sold by several hosts at different rates, and the one we quote changed. The previous rate was still on sale elsewhere. Input and output moved in opposite directions, so the effect on your bill depends on your token mix.
- +4.8% DeepSeek V3.2 (DeepSeek) — blended list price $0.36 to $0.38 per 1M tokens; input +4.1%, output +5.0%. Not a repricing: this model is sold by several hosts at different rates, and the one we quote changed. The previous rate was still on sale elsewhere.
- +2.7% MiniMax M1 (MiniMax) — blended list price $1.66 to $1.71 per 1M tokens; input +37.5%, output 0.0%. Not a repricing: this model is sold by several hosts at different rates, and the one we quote changed. The previous rate was still on sale elsewhere.
- +1.6% Kimi K2.7 Code (Moonshot) — blended list price $2.51 to $2.55 per 1M tokens; input +2.3%, output +1.5%.
Got cheaper (11)
- -67.7% DeepSeek V4.1 Flash (DeepSeek) — blended list price $0.93 to $0.30 per 1M tokens; input -93.3%, output -65.0%.
- -54.3% Kimi K2.6 (Moonshot) — blended list price $3.08 to $1.41 per 1M tokens; input -54.3%, output -54.3%.
- -40.5% DeepSeek V4 Flash 0423 (DeepSeek) — blended list price $0.08 to $0.05 per 1M tokens; input -40.4%, output -40.5%. Not a repricing: this model is sold by several hosts at different rates, and the one we quote changed. The previous rate was still on sale elsewhere.
- -40.0% DeepSeek V4 Pro 0423 (DeepSeek) — blended list price $0.59 to $0.35 per 1M tokens; input -40.0%, output -40.0%. Not a repricing: this model is sold by several hosts at different rates, and the one we quote changed. The previous rate was still on sale elsewhere.
- -37.7% Qwen3 30B A3B Instruct 2507 (Alibaba) — blended list price $0.24 to $0.15 per 1M tokens; input -51.8%, output -35.6%. Not a repricing: this model is sold by several hosts at different rates, and the one we quote changed. The previous rate was still on sale elsewhere.
- -37.5% Hy3 (Tencent) — blended list price $0.41 to $0.26 per 1M tokens; input -37.5%, output -37.5%.
- -30.0% MiniMax M2.7 (MiniMax) — blended list price $0.93 to $0.65 per 1M tokens; input -30.0%, output -30.0%. Not a repricing: this model is sold by several hosts at different rates, and the one we quote changed. The previous rate was still on sale elsewhere.
- -23.1% Qwen3.5-35B-A3B (Alibaba) — blended list price $0.97 to $0.74 per 1M tokens; input -52.0%, output -20.0%. Not a repricing: this model is sold by several hosts at different rates, and the one we quote changed. The previous rate was still on sale elsewhere.
- -20.7% Nemotron 3.5 Lightning (Nvidia) — blended list price $0.16 to $0.13 per 1M tokens; input -25.0%, output -20.0%.
- -10.0% Kimi K3 (Moonshot) — blended list price $11.40 to $10.26 per 1M tokens; input -10.0%, output -10.0%. Not a repricing: this model is sold by several hosts at different rates, and the one we quote changed. The previous rate was still on sale elsewhere.
- -2.0% DeepSeek V4 Flash Vision Exp (DeepSeek) — blended list price $0.53 to $0.52 per 1M tokens; input -2.0%, output -2.0%.
How these figures are calculated
- Prices are blended 30/70 input to output, the same figure the Efficiency score ranks on.
- List prices only. Batch and cache discounts are conditional on the customer implementing them.
- Corrections are excluded rather than annotated: when a published rate was wrong at the source and later fixed, the difference is a data correction, not a price move.
- A change of host is listed but never headlined: most models are sold by several hosts at different rates, this site quotes one of them, and when that host changes the figure moves although nobody repriced anything.
- A missing, zero or negative baseline yields no entry. New listings, renames and delistings get no percentage rather than an invented one.
The full cost-per-request formulas are in the cost per request methodology guide. For what to do about a price move, see the Cloud FinOps skill on GitHub.