GPT-5.6 Luna API Pricing
OpenAI · Budget · Proprietary · released 2026-07. Prices in USD per 1 million tokens. Data snapshot: 2026-09-06.
- Input: $0.20 per 1M tokens
- Output: $1.20 per 1M tokens
- Batch API (input/output): $0.10 / $0.60
- Prompt-cache read: $0.02
- Context window: 1M
- Arena ELO: 1452 (as of 2026-09-01)
- Available on: Direct API, AWS Bedrock, Azure AI Foundry, OpenRouter (curated by hand, verify with the provider)
- AWS Bedrock regions: ap-northeast-1, ap-southeast-1, eu-central-1, eu-west-1, us-east-1, us-east-2, us-west-2
Cost per request by business use case
List price vs optimized (prompt caching on the cached input share, batch API for async workloads, only where the provider publishes those rates).
| Use case | Tokens (in/out) | Cost/req (list) | Cost/req (optimized) | Monthly @100K req |
|---|---|---|---|---|
| Support Ticket | 1,500 / 500 | $0.000900 | $0.000738 | $74 |
| Knowledge Q&A | 2,000 / 800 | $0.0014 | $0.0011 | $111 |
| Meeting Summary | 10,000 / 1,200 | $0.0034 | $0.0016 | $164 |
| Marketing Content | 2,500 / 1,800 | $0.0027 | $0.0026 | $257 |
| Coding Task | 3,000 / 2,000 | $0.0030 | $0.0027 | $273 |
| Invoice Processing | 1,500 / 600 | $0.0010 | $0.000474 | $47 |
| Call Summary | 2,000 / 700 | $0.0012 | $0.000604 | $60 |
| Agent Workflow | 6,000 / 3,000 | $0.0048 | $0.0040 | $404 |
GPT-5.6 Luna pricing by host
A host is a company that runs the model and sells it by the token — the lab that made it, a cloud platform like Amazon Bedrock, Azure or Vertex, or an independent provider. One model, several sellers, and often several prices. This model is sold by 3 different hosts. The price quoted above is the one we publish so that models stay comparable; it is not necessarily the cheapest way to buy this model. List prices as published on 2026-09-06 via OpenRouter — negotiated rates, committed-use discounts and provisioned capacity are not shown.
| Host | Input $/1M | Output $/1M |
|---|---|---|
| OpenAI (the price shown above) | $0.20 | $1.20 |
| Azure (the price shown above) | $0.20 | $1.20 |
| Amazon Bedrock | $0.22 | $1.32 |
Models comparable to GPT-5.6 Luna
The nearest Budget models on Arena ELO, so a substitution is compared on quality as well as price.
- Qwen3.7 Plus (Alibaba) — $0.32 in / $1.28 out per 1M, Arena ELO 1456
- DeepSeek V4 Pro 0423 (DeepSeek) — $0.68 in / $1.35 out per 1M, Arena ELO 1458
- GLM 5 (Zhipu) — $0.60 in / $1.92 out per 1M, Arena ELO 1458
- MiniMax M3 (MiniMax) — $0.30 in / $1.20 out per 1M, Arena ELO 1442
GPT-5.6 Luna head to head
Common questions about GPT-5.6 Luna pricing
- How much does GPT-5.6 Luna cost?
- GPT-5.6 Luna costs $0.20 per million input tokens and $1.20 per million output tokens on the API.
- Does GPT-5.6 Luna offer batch or cached pricing?
- Yes. Batch API pricing is $0.10/M input and $0.60/M output. Prompt-cache reads cost $0.02/M.
- What would GPT-5.6 Luna cost per month?
- For a support-ticket workload (1,500 input + 500 output tokens per request) at 100K requests/month, GPT-5.6 Luna costs about $90 at list prices.
These are list prices: real cost also depends on caching, batching, negotiated discounts, marketplace agreements, and existing contracts. Get a FinOps review from OptimNow or compare all models.
Build the business case for GPT-5.6 Luna in the AI ROI Calculator: add your business value on top of these costs to get ROI, break-even and payback.
All models · Compare side by side · Cloud compute pricing · Price barometer · FinOps guides · API and MCP server