AI API Cost Calculator
Compare pricing across 95 models from 11 providers. Enter your usage below to see exactly what you'd pay — no signup required.
Cost Estimate
Based on published API pricing. Actual costs may vary.
See how prices have changed →Keyboard Shortcuts
How This Calculator Works
Enter your expected usage — tokens per request, requests per day, and days per month — and the calculator shows your estimated cost per request, cost per 1K requests, and monthly total for any of 95 AI models across 11 providers. Use the "Typical request" presets to quickly see what common workloads cost. It also suggests the cheapest alternative if a cheaper model can handle your workload.
Supported Providers
- OpenAI (27 models): GPT-5.6 Sol/Terra/Luna, GPT-5.5, GPT-5.4, GPT-5.3 Codex, GPT-5.2/5.1/5, GPT-5 Pro/nano, GPT-4.1 family, o3/o3 Pro/o4-mini, GPT-oss 120B/20B
- Anthropic (10 models): Opus 5, Fable 5, Sonnet 5, Haiku 4.5, Opus 4.8/4.7, Sonnet 4.6/4.5
- Google (14 models): Gemini 3.1 Pro, Gemini 3.7/3.6/3.5 Flash, Gemini 3 Flash, Gemini 3.5/3.1/2.5 Flash-Lite, Gemini 2.5 Pro/Flash, Nano Banana 2/Pro
- DeepSeek (3 models): V4 Pro, V4 Flash, V3.2
- Mistral (11 models): Large 3, Medium 3.5, Small 4, Codestral, Ministral 3 3B/8B/14B, Magistral Medium/Small (retired)
- Alibaba/Qwen (1 model): Qwen 3.7 Flash (cheapest multimodal)
- Cohere (3 models): Command R+, Command R, Command A
- Meta (Together.ai) (3 models): Llama 4 Scout, Llama 4 Maverick, Llama 3.3 70B
- Moonshot (3 models): Kimi K3, Kimi K2.6, Kimi K2.7 Code
- xAI (5 models): Grok 4.6, Grok 4.5, Grok 4.20, Grok 4.3, Grok Build 0.1
- AI21 (2 models): Jamba 1.7 Large, Jamba Mini
Common Use Cases
Our calculator helps you estimate costs for:
- Chatbots and virtual assistants — typically 500-2000 input tokens, 200-800 output tokens per turn
- Code generation — 1000-5000 input tokens, 500-3000 output tokens per request
- Document analysis — 3000-10000 input tokens, 200-1000 output tokens per document
- Content generation — 500-2000 input tokens, 1000-4000 output tokens per article
Want to compare two models side by side?
Use the Comparison ToolTips to Reduce Your API Costs
The calculator often reveals that a cheaper model can handle your workload. Here are the biggest cost-saving strategies:
- Use budget models for simple tasks: GPT-5.4 nano ($0.20/$1.25) handles 80% of chatbot requests at 1/25th the cost of GPT-5.5
- Optimize prompts: Shorter prompts = fewer input tokens = lower costs
- Set token limits: Don't let models generate 4000 tokens when 200 will do
- Cache responses: Identical prompts can be cached to eliminate redundant API calls
Read our full guide: How to Cut Your AI API Bill in Half: 10 Practical Tips
Get weekly AI pricing updates & cost-saving tips
Join 200+ developers optimizing their AI API spend. No spam, unsubscribe anytime.
Provider Calculators
- Cohere Cost Calculator — Command R+ & Command R pricing
- Moonshot Cost Calculator — Kimi K3 & K2.6 pricing
- Together.ai Cost Calculator — Llama 4 & Llama 3.1 pricing
Related Reading
- AI API Cost Per Request — The metric developers actually need for budgeting LLM costs
- The Complete Guide to LLM Cost Optimization — Strategies to cut your API spend by 40%+
- Cheapest LLM APIs in 2026 — Full ranking of every model by price
- LLM Pricing Cheat Sheet — Quick reference for all 95 models
- How to Estimate Your AI API Costs — Step-by-step cost planning guide
- AI API Cost Comparison Tool — Compare 95 models side by side to find the cheapest option
- May 2026 Pricing Shakeup — Latest price changes across providers
- Cheapest AI API in July 2026 — All 95 models ranked by cost with real scenarios
- AI API Pricing June 2026 — Complete guide to all 95 models, deprecation alerts, migration guide
- AI Startup API Budgets — 5 real budgets from $3/mo to $12K/mo with optimization tips
- Best AI APIs for Structured Output 2026 — JSON mode & function calling compared across 8 models
- Best AI APIs for Chatbots 2026 — All 95 models ranked by cost & quality for conversational AI
- Best AI APIs for RAG 2026 — Embedding + generation models ranked for retrieval-augmented generation
- Best AI APIs for Vision 2026 — Image understanding models ranked by cost & quality
- Best AI Embedding APIs 2026 — Embedding models ranked by MTEB scores & cost
- Best AI Speech APIs 2026 — TTS & STT models ranked by quality & cost
- AI API Cost for Food & Beverage — Restaurant and food production AI budgets
- AI API Cost for Travel & Tourism — Dynamic pricing, chatbots, and recommendation engines