Ranked by effective cost per token across 95 models and 11 providers. Prices in USD per 1 million tokens. See the most expensive โ
Models are ranked by blended cost using a 1:3 input-to-output ratio (typical for chat/assistant workloads). A model costing $0.10/M input + $0.30/M output scores at $1.00/M blended. Context window and provider are noted but don't affect ranking. All prices sourced directly from provider pricing pages.
Price trends: โ Dropped = price decreased recently ยท โ Stable = no change in past 30 days ยท โ Increased = price went up. Trends reflect changes since May 2026.
Alibaba's cheapest multimodal model. 1M context with vision support at the lowest price in the industry.
OpenAI's ultra-budget model. Cheapest OpenAI option at $0.05/M input.
OpenAI's open-source 20B model. Self-host or use via Hugging Face at $0.08/M input.
OpenAI's nano-tier model. 1M context at budget pricing โ ideal for bulk processing.
Google's budget model. 1M context with vision support at $0.10/M input.
Mistral's cheapest edge model. Equal input/output at $0.10/M โ great for high-volume symmetric workloads.
DeepSeek's speed-optimized model. Excellent blended cost at $0.22/M input.
OpenAI's larger open-source model. More capable than the 20B at still-reasonable pricing.
Mistral's efficient small model. Competitive pricing with 128K context.
OpenAI's budget long-context model. 1.05M context at $0.20/M input โ 80% price cut in July 2026.
Use our free calculator to compare all 95 models and find the cheapest option for your specific usage pattern.