๐Ÿ“… Week of July 4, 2026

API Pricing Digest

What changed in AI API pricing โ€” and All tools are now free.

๐Ÿšจ TL;DR โ€” Final Week

โœจ

All Tools Are Free

Price monitoring for 120 models, migration code, cost dashboard โ€” all free, no signup required.

Free Tools โ†’
๐Ÿ”’ No signup required ยท Instant access

๐Ÿ†• New Models This Week

New

GPT-5.4 mini โ€” OpenAI's new value champion

OpenAI launched GPT-5.4 mini at $0.75 input / $4.50 output per 1M tokens. Priced between GPT-4o mini and GPT-4o, it delivers significantly better reasoning than its predecessor while remaining OpenAI's most cost-effective new-gen model. For chatbots, content generation, and data extraction that need stronger reasoning than GPT-4o mini could offer.

New mid-budget option ยท View pricing & alternatives โ†’
New

GPT-5.4 Pro โ€” OpenAI's premium reasoning model

The new flagship from OpenAI. GPT-5.4 Pro targets complex reasoning, code generation, and agentic workflows at $30.00 input / $180.00 output per 1M tokens. Premium pricing positions it alongside Claude Opus 4.8 and GPT-5.5 Pro for the most demanding workloads.

New

Gemini 3.1 Flash-Lite โ€” Google's ultra-cheap option

Google's answer to the price war. Gemini 3.1 Flash-Lite at $0.25 input / $1.50 output per 1M tokens is one of the cheapest multimodal models from a major provider. Perfect for high-volume tasks like classification, routing, and simple Q&A where cost matters most.

โ–ผ Budget tier ยท View pricing & alternatives โ†’
New

DeepSeek V4 Flash โ€” cheapest reasoning-capable model

DeepSeek's latest at $0.22 input / $0.66 output per 1M tokens. Remarkably cheap for a model with strong reasoning capabilities. If you're using GPT-4o for tasks that don't need its full capability, DeepSeek V4 Flash could cut your costs by 90%+.

โ–ผ Budget tier ยท View pricing & alternatives โ†’
New

GPT-oss 120B โ€” OpenAI's open-weight budget play

OpenAI enters the budget open-weight space. GPT-oss 120B at $0.15 input / $0.60 output per 1M tokens. Same price as GPT-4o mini but with a different architecture trade-off. Worth benchmarking against your current model for classification and extraction tasks.

โ–ผ Budget tier ยท View pricing & alternatives โ†’

๐Ÿ“‰ Price Drops

Price Drop

Grok 4.3 โ€” xAI's massive 83% output price cut

xAI rebranded Grok 3 โ†’ Grok 4.3 and slashed pricing from $3.00/$15.00 to $1.25/$2.50 per 1M tokens. That's an 83% output price cut. At $2.50/1M output, Grok 4.3 is now cheaper than Claude Haiku ($5) and competitive with Gemini 3 Flash ($3) for tasks that need solid reasoning. If you dismissed xAI's pricing before, it's time to take another look.

โ–ผ 83% output price cut ยท View pricing & alternatives โ†’

โš ๏ธ Deprecations & Retirements

Deprecation

7 models retired this cycle โ€” check your endpoints

Major cleanup across providers. Anthropic: Claude 4 Opus โ†’ Opus 4.8, Sonnet 4.6 โ†’ Sonnet 5, Sonnet 4 โ†’ Sonnet 4.6. Google: Gemini 2.0 Flash โ†’ 3 Flash, Gemini 2.0 Flash Lite โ†’ 3.1 Flash-Lite. DeepSeek: V3 โ†’ V4 Flash. AI21: Jamba 1.5 โ†’ 1.7. If your code references any of these endpoints, you need to migrate. APIpulse's audit tool now warns about deprecated models and shows migration paths.

Action required if using deprecated endpoints ยท Audit your current model โ†’

๐Ÿ“Š Pricing Trends

Trend

The price floor keeps dropping

Twelve months ago, the cheapest API model from a major provider was ~$0.15/1M input tokens. Today, GPT-oss 20B is at $0.08, Mistral Small at $0.10, and Gemini 2.5 Flash-Lite at $0.10. The "good enough for most tasks" price has nearly halved in a year. If you locked in pricing assumptions 6 months ago, you're overpaying.

โ–ผ ~50% YoY on budget models ยท Full trend analysis โ†’
Trend

Premium models holding steady โ€” for now

While budget models race to the bottom, premium-tier pricing (GPT-5.4 Pro, Claude Opus 4.8, GPT-5.5 Pro) remains at $5-30 input / $25-180 output per 1M tokens. The gap between "cheap" and "best" is now 30-600x. This creates a clear optimization opportunity: route simple tasks to cheap models, reserve premium for complex reasoning.

Are you on a deprecated model? Are you overpaying?

Run a free 30-second audit of your current API model. See exactly how much you could save โ€” and get migration code if you need to switch.

Run My Free Audit โ†’
Takes 30 seconds ยท No signup required ยท See savings instantly

๐Ÿ“ฌ Don't miss next week's changes

Get the API Pricing Digest delivered every Friday. No spam โ€” just pricing intelligence.

Free. Unsubscribe anytime. We respect your inbox.

๐Ÿ“š Past Digests

Browse full archive โ†’

โœจ All tools are free

Stop checking 120 models manually

APIpulse monitors every model across 16 providers. Get alerts when prices drop, migration code ready to paste, and a cost dashboard to track savings. No signup required โ€” everything is free.

Free Tools โ†’
No signup required ยท Instant access