GPT-4o-mini API Pricing
OpenAI · Budget · Proprietary · released 2024-07. Prices in USD per 1 million tokens. Data snapshot: 2026-08-14.
- Input: $0.15 per 1M tokens
- Output: $0.60 per 1M tokens
- Batch API (input/output): $0.08 / $0.30
- Prompt-cache read: $0.07
- Context window: 128K
- Arena ELO: 1219 (as of 2026-08-13)
- Available on: Direct API, Azure AI Foundry, OpenRouter (curated by hand — verify with the provider)
Cost per request by business use case
List price vs optimized (prompt caching on the cached input share, batch API for async workloads — only where the provider publishes those rates).
| Use case | Tokens (in/out) | Cost/req (list) | Cost/req (optimized) | Monthly @100K req |
|---|---|---|---|---|
| Support Ticket | 1,500 / 500 | $0.000525 | $0.000458 | $46 |
| Knowledge Q&A | 2,000 / 800 | $0.000780 | $0.000675 | $68 |
| Meeting Summary | 10,000 / 1,200 | $0.0022 | $0.0012 | $116 |
| Marketing Content | 2,500 / 1,800 | $0.0015 | $0.0014 | $142 |
| Coding Task | 3,000 / 2,000 | $0.0016 | $0.0015 | $154 |
| Invoice Processing | 1,500 / 600 | $0.000585 | $0.000298 | $30 |
| Call Summary | 2,000 / 700 | $0.000720 | $0.000369 | $37 |
| Agent Workflow | 6,000 / 3,000 | $0.0027 | $0.0024 | $239 |
These are list prices — real cost also depends on caching, batching, negotiated discounts, marketplace agreements, and existing contracts. Get a FinOps review from OptimNow or compare all models.