ChatGPT API Pricing 2026: GPT-5.6, GPT-5.5, and Every Model Compared
OpenAI's ChatGPT API ranges from $0.05 to $180 per 1M tokens across 28 active models. The newest additions — GPT-5.6 Sol, Terra, and Luna — launched in July 2026 with 1.05M context windows. Meanwhile, GPT-5.4 nano ($0.20/$1.25) has become the sweet spot for most production workloads: 10x cheaper than GPT-5.5 with 400K context. Here's what every model actually costs and which one fits your budget.
Every OpenAI Model: Pricing Overview
| Model | Input / 1M | Output / 1M | Context | Notes |
|---|---|---|---|---|
| GPT-5.6 Sol | $4.00 | $20.00 | 1.05M | Flagship, frontier reasoning |
| GPT-5.6 Terra | $2.00 | $12.00 | 1.05M | Mid-tier, long context |
| GPT-5.6 Luna | $0.20 | $1.20 | 1.05M | Budget long-context |
| GPT-5.5 | $5.00 | $30.00 | 1.05M | Premium reasoning |
| GPT-5.5 Pro | $30.00 | $180.00 | 1.05M | Maximum capability |
| GPT-5.4 | $2.50 | $15.00 | 400K | Strong general-purpose |
| GPT-5.4 mini | $0.75 | $4.50 | 400K | Mid-range value |
| GPT-5.4 nano | $0.20 | $1.25 | 400K | Best value for production |
| GPT-5.4 Pro | $30.00 | $180.00 | 400K | Premium reasoning |
| GPT-5.3 Codex | $1.75 | $14.00 | 400K | Code-optimized |
| GPT-5.2 | $1.75 | $14.00 | 400K | General-purpose |
| GPT-5.2 Pro | $21.00 | $168.00 | 400K | Premium reasoning |
| GPT-5.1 | $1.25 | $10.00 | 272K | General-purpose |
| GPT-5 | $1.25 | $10.00 | 272K | Sunset Dec 11, 2026 |
| GPT-5 Pro | $15.00 | $120.00 | 272K | Premium reasoning |
| GPT-5 mini | $0.25 | $2.00 | 272K | Sunset Dec 11, 2026 |
| GPT-5 nano | $0.05 | $0.40 | 128K | Cheapest OpenAI model |
| GPT-4.1 | $2.00 | $8.00 | 1M | Long context legacy |
| GPT-4.1 mini | $0.40 | $1.60 | 1M | Budget long context |
| GPT-4.1 nano | $0.10 | $0.40 | 1M | Sunset Oct 23, 2026 |
| o3 | $2.00 | $8.00 | 200K | Reasoning, sunset Dec 11 |
| o3-mini | $1.10 | $4.40 | 200K | Reasoning, sunset Oct 23 |
| o3 Pro | $20.00 | $80.00 | 200K | Premium reasoning, sunset Dec 11 |
| o4-mini | $1.10 | $4.40 | 200K | Reasoning, sunset Oct 23 |
| GPT Audio | $2.50 | $10.00 | 128K | Audio + text |
| GPT Audio Mini | $0.60 | $2.40 | 128K | Budget audio |
| GPT-oss 120B | $0.15 | $0.60 | 128K | Open-source, self-hostable |
| GPT-oss 20B | $0.08 | $0.35 | 128K | Open-source, cheapest |
Note: GPT-4o and GPT-4o mini are deprecated and no longer available for new API accounts. GPT-5 (Dec 11, 2026), GPT-5 mini (Dec 11), o3/o3 Pro (Dec 11), and o3-mini/o4-mini (Oct 23) have confirmed sunset dates. For new projects, use GPT-5.4 nano or GPT-5.6 Luna instead.
Which OpenAI Model Should You Use?
With 28 active models, picking the right one is overwhelming. Here's a quick decision guide:
- Cheapest option: GPT-5 nano ($0.05/$0.40) — classification, extraction, simple formatting
- Best value for production: GPT-5.4 nano ($0.20/$1.25) — chatbots, RAG, summarization with 400K context
- Best code generation: GPT-5.3 Codex ($1.75/$14) — optimized for code tasks
- Strong general-purpose: GPT-5.4 ($2.50/$15) — complex reasoning, 400K context
- Longest context (1.05M): GPT-5.6 Luna ($0.20/$1.20) — budget long-context; GPT-5.6 Sol ($4/$20) — frontier
- Maximum reasoning: GPT-5.5 Pro ($30/$180) — only for tasks that genuinely need it
Real-World Cost Scenarios
Scenario 1: Chatbot (1,000 requests/day)
Average: 800 input tokens, 300 output tokens per request. 30 days/month.
Monthly Chatbot Cost
Key insight: GPT-5.4 nano at $16.20/mo is 39x cheaper than GPT-5.6 Sol for chatbot workloads. For most chatbots, the budget tier is all you need.
Scenario 2: Code Generation (200 requests/day)
Average: 3,000 input tokens, 1,200 output tokens per request. 30 days/month.
Monthly Code Generation Cost
Key insight: For code generation, GPT-5.3 Codex ($648/mo) is optimized for code but costs 50% more than GPT-5.4 ($414/mo). If code quality is paramount, the Codex premium is worth it; otherwise GPT-5.4 handles most code tasks well.
Scenario 3: RAG Pipeline (500 queries/day)
Average: 5,000 input tokens (context + query), 800 output tokens per query. 30 days/month.
Monthly RAG Cost
Scenario 4: Document Summarization (100 documents/day)
Average: 10,000 input tokens, 500 output tokens per document. 30 days/month.
Monthly Summarization Cost
The Hidden Cost: Output Token Multipliers
Output tokens are 4x to 8x more expensive than input tokens across most OpenAI models. This is where costs really add up:
| Model | Input Price | Output Price | Output Multiplier |
|---|---|---|---|
| GPT-5.5 Pro | $30.00 | $180.00 | 6x |
| GPT-5.6 Sol | $5.00 | $30.00 | 6x |
| GPT-5.6 Terra | $2.50 | $15.00 | 6x |
| GPT-5.6 Luna | $1.00 | $6.00 | 6x |
| GPT-5.3 Codex | $1.75 | $14.00 | 8x |
| GPT-5.4 | $2.50 | $15.00 | 6x |
| GPT-5.4 nano | $0.20 | $1.25 | 6.25x |
| GPT-5 nano | $0.05 | $0.40 | 8x |
What this means: Setting max_tokens is critical. An unbounded GPT-5.5 request generating 4,000 output tokens costs $0.12 in output alone — more than 100 GPT-5.4 nano requests.
OpenAI vs Competitors: Which is Cheapest?
| Use Case | Best OpenAI Model | OpenAI Cost | Cheapest Alternative | Alternative Cost |
|---|---|---|---|---|
| Budget Chatbot | GPT-5.4 nano | $16.20/mo | Gemini 3.5 Flash-Lite | $6.00/mo |
| Code Generation | GPT-5.4 | $414.00/mo | DeepSeek V4 Pro | $96.60/mo |
| RAG Pipeline | GPT-5.4 nano | $30.00/mo | Gemini 3.1 Flash-Lite | $12.00/mo |
| Premium Reasoning | GPT-5.6 Sol | $810.00/mo | Claude Opus 5 | $487.50/mo |
| Massive Context | GPT-5.6 Sol (1.05M) | $810.00/mo | Gemini 3.1 Pro (1M) | $487.50/mo |
OpenAI is rarely the cheapest option for any single use case. Its advantage is the broadest model ecosystem — 28 active models from $0.05 to $180/1M tokens — so you can mix and match by workload.
How to Calculate Your ChatGPT API Costs
Cost Formula
Monthly Cost = (Input Tokens × Input Price + Output Tokens × Output Price) × Requests per Month ÷ 1,000,000
Example: 500 requests/day × 2,000 input tokens × $2.50/1M + 500 × 500 output × $15.00/1M = $75 input + $112.50 output = $187.50/month (GPT-5.4)
Or skip the math — use the APIpulse GPT Cost Calculator to compare all OpenAI models side by side with Claude, Gemini, and DeepSeek.
5 Ways to Reduce Your OpenAI API Bill
- Switch to GPT-5.4 nano for most tasks. At $0.20/$1.25, it handles chatbots, RAG, classification, and summarization at a fraction of GPT-5.5's cost. Most workloads don't need frontier reasoning.
- Use GPT-5 nano for simple tasks. Classification, extraction, formatting — GPT-5 nano at $0.05/$0.40 handles these at 99% less cost than GPT-5.5.
- Set max_tokens religiously. Output tokens cost 4-8x more than input. Setting max_tokens to 500 instead of leaving it unbounded can cut costs 60%.
- Implement model routing. Route simple queries to GPT-5 nano, moderate to GPT-5.4 nano, and only complex reasoning to GPT-5.6 Sol. This can reduce costs 40-70%.
- Consider GPT-oss models. GPT-oss 120B ($0.15/$0.60) and GPT-oss 20B ($0.08/$0.35) are open-source models good for batch workloads and self-hosted deployments.
The Bottom Line
Many developers overpay for OpenAI API. Test the lowest-cost model that meets your measured quality and latency requirements. GPT-5.4 nano ($0.20/$1.25) fits compact, high-volume tasks, while GPT-5.6 Luna ($0.20/$1.20) adds a 1.05M context window at a similar standard token price. Reserve GPT-5.6 Terra or Sol for workloads where evaluation results justify the higher cost.
Calculate your exact ChatGPT API costs. Enter your usage and compare with every alternative.
Try the Free GPT Calculator or Compare All Models🎯 Rate Your API Setup in 30 Seconds
Get an A+ to F grade on your AI API costs. See how you compare and find cheaper alternatives instantly.
Get Your Cost Score →