Mistral Small 4 API Pricing: Europe's Best Budget Model at $0.15/M
Mistral Small 4 offers capable general-purpose AI at $0.15/M input, $0.60/M output — the cheapest general-purpose model from a European provider with GDPR-friendly infrastructure.
TL;DR
- Price: $0.15/M input, $0.60/M output — 4:1 output ratio
- Context: 128K tokens — standard for budget models
- Provider: Mistral (Paris, France) — European data sovereignty
- Best for: Chatbots, code assistance, document analysis, EU compliance needs
- vs GPT-5 nano: 3x more expensive but more capable reasoning
- vs Mistral Large 3: 3.3x cheaper, same provider, slightly smaller context
Mistral Small 4 Pricing Breakdown
At $0.15 per million input tokens and $0.60 per million output tokens, Mistral Small 4 sits in the budget tier — more capable than ultra-budget models (Qwen 3.7 Flash, GPT-5 nano) but significantly cheaper than mid-tier options (GPT-5.4, Claude Sonnet 5). Here's how costs scale:
| Monthly Volume | Input Cost | Output Cost | Total (50/50 split) |
|---|---|---|---|
| 1M tokens | $0.15 | $0.60 | $0.38 |
| 10M tokens | $1.50 | $6.00 | $3.75 |
| 100M tokens | $15.00 | $60.00 | $37.50 |
| 1B tokens | $150.00 | $600.00 | $375.00 |
At 100M tokens/month, Mistral Small 4 costs $37.50 — about half the cost of Mistral Large 3 ($100) and a fraction of Claude Sonnet 5 ($600) or GPT-5.4 ($87.50).
Mistral Small 4 vs Other Budget Models
| Model | Input $/M | Output $/M | Context | Provider | Region |
|---|---|---|---|---|---|
| Qwen 3.7 Flash | $0.03 | $0.13 | 1M | Alibaba | China |
| GPT-5 nano | $0.05 | $0.40 | 128K | OpenAI | US |
| Gemini 2.5 Flash-Lite | $0.10 | $0.40 | 1M | US | |
| Ministral 3 3B | $0.10 | $0.10 | 128K | Mistral | EU |
| Mistral Small 4 | $0.15 | $0.60 | 128K | Mistral | EU |
| DeepSeek V4 Flash | $0.14 | $0.28 | 1M | DeepSeek | China |
| Devstral Small 2 | $0.10 | $0.30 | 128K | Mistral | EU |
| Mistral Large 3 | $0.50 | $1.50 | 262K | Mistral | EU |
Mistral Small 4 is slightly more expensive than DeepSeek V4 Flash ($0.14/$0.28) and Ministral 3 3B ($0.10/$0.10), but it's a larger, more capable model. The key differentiator is European hosting — if you need EU data sovereignty, Mistral Small 4 is the cheapest general-purpose option.
Mistral Model Family: Which One?
| Model | Input $/M | Output $/M | Context | Best For |
|---|---|---|---|---|
| Ministral 3 3B | $0.10 | $0.10 | 128K | Ultra-budget, edge, output-heavy |
| Devstral Small 2 | $0.10 | $0.30 | 128K | Budget coding tasks |
| Mistral Small 4 | $0.15 | $0.60 | 128K | General-purpose budget |
| Magistral Small | $0.50 | $1.50 | 256K | Reasoning tasks |
| Codestral | $0.30 | $0.90 | 256K | Code generation specialist |
| Mistral Medium 3.5 | $1.50 | $7.50 | 128K | Complex reasoning |
| Mistral Large 3 | $0.50 | $1.50 | 262K | Best Mistral quality |
Decision guide: Need the absolute cheapest? Ministral 3 3B. Need coding? Devstral Small 2 or Codestral. Need general-purpose quality on a budget? Mistral Small 4. Need the best Mistral quality? Mistral Large 3.
Best Use Cases for Mistral Small 4
🇪🇺 EU-Compliant Chatbots
Customer support and internal tools that need European data residency. Mistral's Paris-based infrastructure keeps data in the EU.
💻 Code Assistance
Code completion, review, and explanation at budget pricing. Good for developer tools that need quality without premium costs.
📄 Document Processing
Summarization, extraction, and analysis of business documents. 128K context handles most document lengths.
🤖 AI Agent Workflows
Multi-step agent chains where each step needs decent quality but cost control is important. Good price-to-capability ratio.
🎓 Education & Research
Research tools and educational platforms that need reliable language understanding at affordable per-token costs.
🔄 High-Volume Classification
Content moderation, intent detection, and sentiment analysis at scale. More capable than ultra-budget models for nuanced tasks.
Real-World Cost: 50K Support Conversations/Month
Consider a customer support chatbot handling 50,000 conversations per month. Each conversation averages 2,000 input tokens (conversation history + knowledge base context) and 500 output tokens (response).
| Model | Input Cost | Output Cost | Total/Month |
|---|---|---|---|
| Qwen 3.7 Flash | $3.00 | $3.25 | $6.25 |
| GPT-5 nano | $5.00 | $10.00 | $15.00 |
| DeepSeek V4 Flash | $14.00 | $7.00 | $21.00 |
| Mistral Small 4 | $15.00 | $15.00 | $30.00 |
| Mistral Large 3 | $50.00 | $37.50 | $87.50 |
| Claude Sonnet 5 | $200.00 | $250.00 | $450.00 |
At $30/month for 50K conversations, Mistral Small 4 is 60% cheaper than Mistral Large 3 and 93% cheaper than Claude Sonnet 5 — while offering solid general-purpose quality for support use cases.
Compare Mistral Small 4 Pricing
See how Mistral Small 4 stacks up against 94 models across 11 providers.
Frequently Asked Questions
Related Pages
- Ministral 3 3B Pricing — Mistral's cheapest model at $0.10/$0.10
- DeepSeek V4 Pricing — China's cheapest budget model at $0.14/$0.28
- GPT-5 nano Pricing — OpenAI's cheapest model at $0.05/$0.40
- Top 10 Cheapest LLM APIs — Ranked by price per million tokens
- API Cost Calculator — Compare costs across 94 models