July 18, 2026 · 6 min read

15 New AI Models Added: GPT-5.6 Pro, Claude Fast, Grok 4.20 & More

The APIpulse database just got a major update — 15 new models from OpenAI, Anthropic, Google, xAI, and Mistral. We now track 85 models across 10 providers. Here's what's new, what it costs, and when to use each.

OpenAI: GPT-5.6 Pro Variants + Audio Models

OpenAI added Pro variants for the entire GPT-5.6 lineup, plus their first audio-specific API models.

GPT-5.6 Luna Pro
$1.00 / $6.00 per 1M tokens
1.05M context · 128K output
Pro Tier
GPT-5.6 Terra Pro
$2.50 / $15.00 per 1M tokens
1.05M context · 128K output
Pro Tier
GPT-5.6 Sol Pro
$5.00 / $30.00 per 1M tokens
1.05M context · 128K output
Pro Tier
GPT Audio
$2.50 / $10.00 per 1M tokens
128K context · Audio I/O
Audio
GPT Audio Mini
$0.60 / $2.40 per 1M tokens
128K context · Audio I/O
Audio
o3 Pro
$20.00 / $80.00 per 1M tokens
200K context · Reasoning
Reasoning

Key insight: The GPT-5.6 Pro variants have the same per-token price as their non-Pro counterparts. The "Pro" designation likely refers to higher rate limits or priority access rather than different model weights. GPT Audio at $2.50/$10 is OpenAI's first dedicated audio API — useful for voice agents and transcription pipelines. o3 Pro at $20/$80 is the most expensive reasoning model available, competing with Claude Opus on complex tasks.

Anthropic: Claude Fast Speed Tiers

Anthropic introduced "Fast" variants for their flagship models — same capabilities, significantly higher throughput, premium pricing.

Claude Opus 4.8 Fast
$10.00 / $50.00 per 1M tokens
1M context · 2× price for 3× speed
Speed Tier
Claude Opus 4.7 Fast
$30.00 / $150.00 per 1M tokens
1M context · 6× price for speed
Speed Tier

Key insight: Claude Opus 4.8 Fast at $10/$50 is 2× the standard Opus 4.8 price. Claude Opus 4.7 Fast at $30/$150 is a massive premium — 6× the standard price. These are for latency-sensitive production workloads where response time matters more than cost. If you're building a real-time chat interface or code completion tool, Fast tiers are worth it. For batch processing, stick with standard.

xAI: Grok 4.20 and Grok 4.3

Grok 4.20
$1.20 / $2.50 per 1M tokens
2M context — largest available
New
Grok 4.3
$1.20 / $2.50 per 1M tokens
1M context
New

Key insight: Grok 4.20 has a massive 2M token context window — the largest of any model we track. At $1.20/$2.50, it's priced competitively with Claude Sonnet 5 ($2/$10) but with 2× the context. If you're working with extremely long documents, codebases, or multi-turn conversations, Grok 4.20 is the best value for context-heavy work.

Mistral: Large 3 and Small 4

Mistral Large 3
$0.50 / $1.50 per 1M tokens
262K context
New
Mistral Small 4
$0.15 / $0.60 per 1M tokens
262K context
New

Key insight: Mistral Large 3 at $0.50/$1.50 is 80% cheaper than Claude Sonnet 5 ($2/$10) on input. Mistral Small 4 at $0.15/$0.60 is one of the cheapest capable models available — great for high-volume classification, extraction, and routing tasks. Both support 262K context. For EU-based teams needing GDPR compliance, Mistral remains the best European option.

Google: Nano Banana Image Models

Nano Banana 2 (Gemini 3.1 Flash Image)
$0.50 / $3.00 per 1M tokens
131K context · Image generation
Image Gen
Nano Banana Pro (Gemini 3 Pro Image)
$2.00 / $12.00 per 1M tokens
65K context · Image generation
Image Gen

Key insight: Google's "Nano Banana" models are their image generation API, built on Gemini architecture. Nano Banana 2 at $0.50/$3.00 is the cheapest image generation option. Nano Banana Pro at $2/$12 offers higher quality. These compete with DALL-E and Midjourney on the API side.

Quick Comparison: New Models vs Existing Best Options

Model Input / Output per 1M Context
Mistral Small 4 $0.15 / $0.60 262K
GPT-5.6 Luna Pro $1.00 / $6.00 1.05M
Grok 4.20 $1.20 / $2.50 2M
Claude Opus 4.8 Fast $10.00 / $50.00 1M
Claude Opus 4.7 Fast $30.00 / $150.00 1M
o3 Pro $20.00 / $80.00 200K

Which One Should You Use?

Cheapest capable model

Mistral Small 4 — $0.15/$0.60 with 262K context

Largest context window

Grok 4.20 — 2M tokens at $1.20/$2.50

Fastest responses

Claude Opus 4.8 Fast — premium speed for latency-sensitive apps

Voice/audio apps

GPT Audio Mini — $0.60/$2.40 for audio I/O

Complex reasoning

o3 Pro — $20/$80 for the hardest problems

EU compliance

Mistral Large 3 — $0.50/$1.50, GDPR-native

The Bigger Picture

These 15 models push the APIpulse database to 85 models across 10 providers. The trends are clear:

The best strategy in 2026 isn't picking one model — it's using a multi-model approach. Use Mistral Small 4 for high-volume cheap tasks, Grok 4.20 for long-context work, Claude Opus 4.8 Fast for real-time apps, and o3 Pro for complex reasoning. Our decision tree can help you pick.

Compare all 85 models side by side

Open Cost Calculator →

Pricing data verified Jul 18, 2026. Prices per 1M tokens. See our full pricing table for all 85 models.

🎯 Rate Your API Setup in 30 Seconds

Get an A+ to F grade on your AI API costs. See how you compare and find cheaper alternatives instantly.

Get Your Cost Score →

📊 Generate Your Personalized API Cost Report

Select your model, enter your monthly spend, and get a custom savings report with cheaper alternatives — free, in 60 seconds.

Want to optimize your AI API costs?

APIpulse includes free cost comparisons, exports, and recommendations that can save you up to 40%.

Free Cost Audit →
🔧 Free Embeddable Pricing Widget
Add live AI API pricing to your docs, blog, or README with one script tag. 85 models, auto-updating.
Get the Free Widget → Free MCP Server →