Ling 3.0 Flash Sante API pricing
inclusionai · Budget · released 2026-09. Prices in USD per 1 million tokens. Data snapshot: 2026-10-09.
Ling 3.0 Flash Sante has held the same list price since 2026-10-09 and is the 9th cheapest of 144 Budget models. Only one host, Novita, sells it, so there is no second price to shop. No Arena score is published for it, so it carries no Efficiency score and cannot hold the FinOps Friendly badge.
- Input: $0.042 per 1M tokens
- Output: $0.1232 per 1M tokens
- Batch API (in/out): Not published
- Prompt-cache read: $0.0084 per 1M tokens
- Context window: 262K tokens
- Arena ELO: Not published
Market position
On blended price ($0.0988 per 1M tokens at 30% input, 70% output) it is cheaper than 94% of Budget models: 9th cheapest of 144, tied with one other. Across every category it is cheaper than 97% of priced text models in the catalogue: 9th cheapest of 286, tied with one other.
Licence, size and capabilities
No licence is recorded for Ling 3.0 Flash Sante, so this page does not say whether it can be self-hosted. Its published size is 5.1B parameters. Listed capabilities: Text, Reasoning, Agents. Filling its 262K context window once costs $0.011 in input tokens at list price.
Pricing by host
A host is a company that runs the model and sells it by the token: the model maker, a cloud platform like Amazon Bedrock, Azure or Vertex, or an independent host. One model, several hosts, and often several prices. This model is sold by 1 host. The price we quote elsewhere on this page is marked. We quote one host's rate so that models stay comparable; it is not a recommendation to buy there.
Novita is the only host with a published rate, $0.042 input / $0.1232 output per 1M tokens, which is the price we quote.
| Host | Input $/1M | Output $/1M |
|---|---|---|
| Novita (price we quote) | $0.042 | $0.1232 |
List prices as published on 2026-10-09, via OpenRouter. Negotiated rates, committed-use discounts and provisioned capacity are not shown.
Monthly cost at 100K requests, by business use case
From the after-discounts figure (prompt caching on the cached input share, batch API for async workloads, only where inclusionai publishes those rates), which assumes you implement both, to the list price. One figure means no discount is published.
| Use case | Tokens (in/out) | Monthly (after discounts – list) |
|---|---|---|
| Support Ticket | 1,500 / 500 | $9 – $12 |
| Knowledge Q&A | 2,000 / 800 | $14 – $18 |
| Meeting Summary | 10,000 / 1,200 | $53 – $57 |
| Marketing Content | 2,500 / 1,800 | $31 – $33 |
| Coding Task | 3,000 / 2,000 | $32 – $37 |
| Invoice Processing | 1,500 / 600 | $12 – $14 |
| Call Summary | 2,000 / 700 | $16 – $17 |
| Agent Workflow | 6,000 / 3,000 | $48 – $62 |
Price history
Ling 3.0 Flash Sante entered our daily price archive on 2026-10-09; tracking began on 2026-08-15. It listed at $0.042 input / $0.1232 output per 1M tokens that day. No list price change since 2026-10-09. Last checked against the 2026-10-09 snapshot.
Alternatives to Ling 3.0 Flash Sante
- Next cheaper Budget model: Gemma 3 4B (Google) at $0.085 blended, 14% less.
- Next more expensive: Qwen3.7 Flash (Alibaba) at $0.10 blended, 1% more.
Other models from inclusionAI
inclusionAI has 3 other models in the catalogue. Prices per 1M tokens, input then output.
- Ling 3.0 Flash VL: $0.021 / $0.0616
- Ling 3.0 Flash: $0.021 / $0.063
- Ling 3.0 Flash Fin: $0.042 / $0.1232
All 4 inclusionAI models and prices
Common questions about Ling 3.0 Flash Sante pricing
- How much does Ling 3.0 Flash Sante cost?
- Ling 3.0 Flash Sante costs $0.042 per 1M input tokens and $0.1232 per 1M output tokens at Novita, the host whose rate we quote. Prices as listed on 2026-10-09.
- Which host is cheapest for Ling 3.0 Flash Sante?
- Novita is the only host with a published rate, $0.042 input / $0.1232 output per 1M tokens, which is the price we quote. Host rates as listed on 2026-10-09.
- Does Ling 3.0 Flash Sante offer batch or cached pricing?
- Yes, for prompt caching only. Cache reads cost $0.0084 per 1M tokens, as listed on 2026-10-09. No batch pricing is published.
- What would Ling 3.0 Flash Sante cost per month?
- With the Support Ticket profile (1,500 input and 500 output tokens per request), 100,000 requests a month cost about $12 on Ling 3.0 Flash Sante at the list price quoted on 2026-10-09.
- Can Ling 3.0 Flash Sante be self-hosted?
- No licence is recorded for Ling 3.0 Flash Sante, so this page does not say whether it can be self-hosted.
- Has the price of Ling 3.0 Flash Sante changed?
- Ling 3.0 Flash Sante entered our daily price archive on 2026-10-09; tracking began on 2026-08-15. It listed at $0.042 input / $0.1232 output per 1M tokens that day. No list price change since 2026-10-09. Last checked against the 2026-10-09 snapshot.
These are list prices. Your real cost also depends on caching and batching, negotiated discounts, marketplace agreements, and the contracts you already have. OptimNow audits exactly that.
Build the business case · Get a FinOps review
LLM pricing · Models by maker · Cloud compute pricing · Compare models · Price barometer · FinOps guides · API and MCP server