ChatGPT API Pricing 2026: GPT-5.6, GPT-5.5, and Every Model Compared

OpenAI's ChatGPT API ranges from $0.05 to $180 per 1M tokens across 28 active models. The newest additions — GPT-5.6 Sol, Terra, and Luna — launched in July 2026 with 1.05M context windows. Meanwhile, GPT-5.4 nano ($0.20/$1.25) has become the sweet spot for most production workloads: 10x cheaper than GPT-5.5 with 400K context. Here's what every model actually costs and which one fits your budget.

Every OpenAI Model: Pricing Overview

Model Input / 1M Output / 1M Context Notes
GPT-5.6 Sol$4.00$20.001.05MFlagship, frontier reasoning
GPT-5.6 Terra$2.00$12.001.05MMid-tier, long context
GPT-5.6 Luna$0.20$1.201.05MBudget long-context
GPT-5.5$5.00$30.001.05MPremium reasoning
GPT-5.5 Pro$30.00$180.001.05MMaximum capability
GPT-5.4$2.50$15.00400KStrong general-purpose
GPT-5.4 mini$0.75$4.50400KMid-range value
GPT-5.4 nano$0.20$1.25400KBest value for production
GPT-5.4 Pro$30.00$180.00400KPremium reasoning
GPT-5.3 Codex$1.75$14.00400KCode-optimized
GPT-5.2$1.75$14.00400KGeneral-purpose
GPT-5.2 Pro$21.00$168.00400KPremium reasoning
GPT-5.1$1.25$10.00272KGeneral-purpose
GPT-5$1.25$10.00272KSunset Dec 11, 2026
GPT-5 Pro$15.00$120.00272KPremium reasoning
GPT-5 mini$0.25$2.00272KSunset Dec 11, 2026
GPT-5 nano$0.05$0.40128KCheapest OpenAI model
GPT-4.1$2.00$8.001MLong context legacy
GPT-4.1 mini$0.40$1.601MBudget long context
GPT-4.1 nano$0.10$0.401MSunset Oct 23, 2026
o3$2.00$8.00200KReasoning, sunset Dec 11
o3-mini$1.10$4.40200KReasoning, sunset Oct 23
o3 Pro$20.00$80.00200KPremium reasoning, sunset Dec 11
o4-mini$1.10$4.40200KReasoning, sunset Oct 23
GPT Audio$2.50$10.00128KAudio + text
GPT Audio Mini$0.60$2.40128KBudget audio
GPT-oss 120B$0.15$0.60128KOpen-source, self-hostable
GPT-oss 20B$0.08$0.35128KOpen-source, cheapest

Note: GPT-4o and GPT-4o mini are deprecated and no longer available for new API accounts. GPT-5 (Dec 11, 2026), GPT-5 mini (Dec 11), o3/o3 Pro (Dec 11), and o3-mini/o4-mini (Oct 23) have confirmed sunset dates. For new projects, use GPT-5.4 nano or GPT-5.6 Luna instead.

Which OpenAI Model Should You Use?

With 28 active models, picking the right one is overwhelming. Here's a quick decision guide:

Real-World Cost Scenarios

Scenario 1: Chatbot (1,000 requests/day)

Average: 800 input tokens, 300 output tokens per request. 30 days/month.

Monthly Chatbot Cost

GPT-5.6 Sol $630.00/mo
GPT-5.4 $195.00/mo
GPT-5.4 nano $16.20/mo
GPT-5 nano $4.80/mo
GPT-oss 20B $4.02/mo

Key insight: GPT-5.4 nano at $16.20/mo is 39x cheaper than GPT-5.6 Sol for chatbot workloads. For most chatbots, the budget tier is all you need.

Scenario 2: Code Generation (200 requests/day)

Average: 3,000 input tokens, 1,200 output tokens per request. 30 days/month.

Monthly Code Generation Cost

GPT-5.5 $2,250.00/mo
GPT-5.3 Codex $648.00/mo
GPT-5.4 $414.00/mo
GPT-5.4 mini $162.00/mo
GPT-5.4 nano $48.60/mo

Key insight: For code generation, GPT-5.3 Codex ($648/mo) is optimized for code but costs 50% more than GPT-5.4 ($414/mo). If code quality is paramount, the Codex premium is worth it; otherwise GPT-5.4 handles most code tasks well.

Scenario 3: RAG Pipeline (500 queries/day)

Average: 5,000 input tokens (context + query), 800 output tokens per query. 30 days/month.

Monthly RAG Cost

GPT-5.6 Sol $810.00/mo
GPT-5.4 $285.00/mo
GPT-5.4 mini $103.50/mo
GPT-5.4 nano $30.00/mo
GPT-5 nano $9.00/mo

Scenario 4: Document Summarization (100 documents/day)

Average: 10,000 input tokens, 500 output tokens per document. 30 days/month.

Monthly Summarization Cost

GPT-5.6 Sol $600.00/mo
GPT-5.4 $187.50/mo
GPT-5.4 nano $22.50/mo
GPT-5 nano $6.60/mo

The Hidden Cost: Output Token Multipliers

Output tokens are 4x to 8x more expensive than input tokens across most OpenAI models. This is where costs really add up:

Model Input Price Output Price Output Multiplier
GPT-5.5 Pro $30.00 $180.00 6x
GPT-5.6 Sol $5.00 $30.00 6x
GPT-5.6 Terra $2.50 $15.00 6x
GPT-5.6 Luna $1.00 $6.00 6x
GPT-5.3 Codex $1.75 $14.00 8x
GPT-5.4 $2.50 $15.00 6x
GPT-5.4 nano $0.20 $1.25 6.25x
GPT-5 nano $0.05 $0.40 8x

What this means: Setting max_tokens is critical. An unbounded GPT-5.5 request generating 4,000 output tokens costs $0.12 in output alone — more than 100 GPT-5.4 nano requests.

OpenAI vs Competitors: Which is Cheapest?

Use Case Best OpenAI Model OpenAI Cost Cheapest Alternative Alternative Cost
Budget Chatbot GPT-5.4 nano $16.20/mo Gemini 3.5 Flash-Lite $6.00/mo
Code Generation GPT-5.4 $414.00/mo DeepSeek V4 Pro $96.60/mo
RAG Pipeline GPT-5.4 nano $30.00/mo Gemini 3.1 Flash-Lite $12.00/mo
Premium Reasoning GPT-5.6 Sol $810.00/mo Claude Opus 5 $487.50/mo
Massive Context GPT-5.6 Sol (1.05M) $810.00/mo Gemini 3.1 Pro (1M) $487.50/mo

OpenAI is rarely the cheapest option for any single use case. Its advantage is the broadest model ecosystem — 28 active models from $0.05 to $180/1M tokens — so you can mix and match by workload.

How to Calculate Your ChatGPT API Costs

Cost Formula

Monthly Cost = (Input Tokens × Input Price + Output Tokens × Output Price) × Requests per Month ÷ 1,000,000

Example: 500 requests/day × 2,000 input tokens × $2.50/1M + 500 × 500 output × $15.00/1M = $75 input + $112.50 output = $187.50/month (GPT-5.4)

Or skip the math — use the APIpulse GPT Cost Calculator to compare all OpenAI models side by side with Claude, Gemini, and DeepSeek.

5 Ways to Reduce Your OpenAI API Bill

  1. Switch to GPT-5.4 nano for most tasks. At $0.20/$1.25, it handles chatbots, RAG, classification, and summarization at a fraction of GPT-5.5's cost. Most workloads don't need frontier reasoning.
  2. Use GPT-5 nano for simple tasks. Classification, extraction, formatting — GPT-5 nano at $0.05/$0.40 handles these at 99% less cost than GPT-5.5.
  3. Set max_tokens religiously. Output tokens cost 4-8x more than input. Setting max_tokens to 500 instead of leaving it unbounded can cut costs 60%.
  4. Implement model routing. Route simple queries to GPT-5 nano, moderate to GPT-5.4 nano, and only complex reasoning to GPT-5.6 Sol. This can reduce costs 40-70%.
  5. Consider GPT-oss models. GPT-oss 120B ($0.15/$0.60) and GPT-oss 20B ($0.08/$0.35) are open-source models good for batch workloads and self-hosted deployments.

The Bottom Line

Many developers overpay for OpenAI API. Test the lowest-cost model that meets your measured quality and latency requirements. GPT-5.4 nano ($0.20/$1.25) fits compact, high-volume tasks, while GPT-5.6 Luna ($0.20/$1.20) adds a 1.05M context window at a similar standard token price. Reserve GPT-5.6 Terra or Sol for workloads where evaluation results justify the higher cost.

Calculate your exact ChatGPT API costs. Enter your usage and compare with every alternative.

Try the Free GPT Calculator or Compare All Models

🎯 Rate Your API Setup in 30 Seconds

Get an A+ to F grade on your AI API costs. See how you compare and find cheaper alternatives instantly.

Get Your Cost Score →

Found this useful? Share it:

Save money: 📊 Live API Pricing · Cost Optimizer — find out how much you could save by switching models. Free tool.

🔧 Free Embeddable Pricing Widget
Add live AI API pricing to your docs, blog, or README with one script tag. 95 models, auto-updating.
Get the Free Widget → Free MCP Server →