Qwen3.5-122B-A10B API Pricing
Alibaba · Mid-tier · Qwen · released 2026-03. Prices in USD per 1 million tokens. Data snapshot: 2026-09-06.
- Input: $0.29 per 1M tokens
- Output: $2.40 per 1M tokens
- Batch API (input/output): not published
- Prompt-cache read: not published
- Context window: 262K
- Arena ELO: 1417 (as of 2026-09-01)
Cost per request by business use case
List price vs optimized (prompt caching on the cached input share, batch API for async workloads, only where the provider publishes those rates).
| Use case | Tokens (in/out) | Cost/req (list) | Cost/req (optimized) | Monthly @100K req |
|---|---|---|---|---|
| Support Ticket | 1,500 / 500 | $0.0016 | $0.0016 | $164 |
| Knowledge Q&A | 2,000 / 800 | $0.0025 | $0.0025 | $250 |
| Meeting Summary | 10,000 / 1,200 | $0.0058 | $0.0058 | $578 |
| Marketing Content | 2,500 / 1,800 | $0.0050 | $0.0050 | $505 |
| Coding Task | 3,000 / 2,000 | $0.0057 | $0.0057 | $567 |
| Invoice Processing | 1,500 / 600 | $0.0019 | $0.0019 | $188 |
| Call Summary | 2,000 / 700 | $0.0023 | $0.0023 | $226 |
| Agent Workflow | 6,000 / 3,000 | $0.0089 | $0.0089 | $894 |
Qwen3.5-122B-A10B pricing by host
A host is a company that runs the model and sells it by the token — the lab that made it, a cloud platform like Amazon Bedrock, Azure or Vertex, or an independent provider. One model, several sellers, and often several prices. This model is sold by 5 different hosts. The price quoted above is the one we publish so that models stay comparable; it is not necessarily the cheapest way to buy this model. List prices as published on 2026-09-06 via OpenRouter — negotiated rates, committed-use discounts and provisioned capacity are not shown. These hosts do not all serve the model at the same numeric precision, so the cheapest is not a like-for-like substitute for the dearest.
| Host | Input $/1M | Output $/1M |
|---|---|---|
| SiliconFlow — fp8 | $0.26 | $2.08 |
| Alibaba | $0.26 | $2.08 |
| DeepInfra (the price shown above) — fp4 | $0.29 | $2.40 |
| AtlasCloud — fp8 | $0.30 | $2.40 |
| Novita — bf16 | $0.40 | $3.20 |
Models comparable to Qwen3.5-122B-A10B
The nearest Mid-tier models on Arena ELO, so a substitution is compared on quality as well as price.
- Kimi K2 0711 (Moonshot) — $0.57 in / $2.30 out per 1M, Arena ELO 1418
- GPT-4.1 (OpenAI) — $2.00 in / $8.00 out per 1M, Arena ELO 1414
- Mistral Large (Mistral) — $2.00 in / $6.00 out per 1M, Arena ELO 1414
- Claude Haiku 4.5 (Anthropic) — $1.00 in / $5.00 out per 1M, Arena ELO 1413
Common questions about Qwen3.5-122B-A10B pricing
- How much does Qwen3.5-122B-A10B cost?
- Qwen3.5-122B-A10B costs $0.29 per million input tokens and $2.40 per million output tokens on the API.
- Does Qwen3.5-122B-A10B offer batch or cached pricing?
- No batch or cache pricing is published for this model.
- What would Qwen3.5-122B-A10B cost per month?
- For a support-ticket workload (1,500 input + 500 output tokens per request) at 100K requests/month, Qwen3.5-122B-A10B costs about $164 at list prices.
These are list prices: real cost also depends on caching, batching, negotiated discounts, marketplace agreements, and existing contracts. Get a FinOps review from OptimNow or compare all models.
Build the business case for Qwen3.5-122B-A10B in the AI ROI Calculator: add your business value on top of these costs to get ROI, break-even and payback.
All models · Compare side by side · Cloud compute pricing · Price barometer · FinOps guides · API and MCP server