Gemini 3.5 Flash-Lite API Pricing: Google's Thinking Budget Model at $0.30/M Tokens

Gemini 3.5 Flash-Lite is Google's cheapest thinking model — built-in reasoning at $0.30/$2.50 with 1M context, vision, and no hidden thinking token charges.

Updated Aug 10, 2026 · 93 models tracked across 11 providers

TL;DR

Gemini 3.5 Flash-Lite Pricing Breakdown

At $0.30 per million input tokens and $2.50 per million output tokens, Gemini 3.5 Flash-Lite is Google's cheapest model with thinking capability. The output price includes all reasoning tokens, so you won't get surprise charges from hidden chain-of-thought steps.

Monthly Volume Input Cost Output Cost Total Cost
1M tokens$0.30$2.50$2.80
10M tokens$3.00$25.00$28.00
100M tokens$30.00$250.00$280.00
1B tokens$300.00$2,500.00$2,800.00

At 100M tokens/month, Gemini 3.5 Flash-Lite costs $280. The same volume on Gemini 3.1 Flash-Lite (no thinking) would cost $175, and on Gemini 3 Flash it would cost $350.

How Gemini 3.5 Flash-Lite Compares to Other Budget Models

Model Input $/M Output $/M Context Vision Thinking Provider
Qwen 3.7 Flash $0.03 $0.13 1M Alibaba
DeepSeek V4 Flash $0.14 $0.28 1M DeepSeek
GPT-5.4 nano $0.20 $1.25 400K OpenAI
Gemini 3.1 Flash-Lite $0.25 $1.50 1M Google
Gemini 3.5 Flash-Lite $0.30 $2.50 1M Google
Gemini 3 Flash $0.50 $3.00 1M Google
GPT-5.4 mini $0.75 $4.50 400K OpenAI
Claude Haiku 4.5 $1.00 $5.00 200K Anthropic

Gemini 3.5 Flash-Lite is the only sub-$1 model with both thinking capability and 1M context. For tasks that need step-by-step reasoning (data analysis, nuanced classification, document Q&A), it offers the best value in the budget tier.

Google Flash-Lite Family: Which One Should You Choose?

ModelInput $/MOutput $/MThinkingBest For
Gemini 2.5 Flash-Lite$0.10$0.40Cheapest Google option — simple extraction, classification
Gemini 3.1 Flash-Lite$0.25$1.50Budget multimodal — faster, no thinking overhead
Gemini 3.5 Flash-Lite$0.30$2.50Budget reasoning — step-by-step logic, analysis

Rule of thumb: Use Gemini 2.5 Flash-Lite for simple, high-volume tasks. Use Gemini 3.1 Flash-Lite when you need better quality but no reasoning. Use Gemini 3.5 Flash-Lite when tasks need step-by-step thinking — the 20% price premium over 3.1 often pays for itself in output quality.

When to Use Gemini 3.5 Flash-Lite

📊 Data Analysis

Analyze datasets, identify patterns, generate insights. Thinking capability handles multi-step analysis without upgrading to a premium model.

💻 Code Generation

Generate and debug code with step-by-step reasoning. Catches edge cases that non-thinking models miss, at budget pricing.

📄 Document Q&A

Answer complex questions about long documents. 1M context handles entire codebases or legal contracts; thinking improves answer quality.

🏷️ Nuanced Classification

Classification tasks that need reasoning — intent detection, sentiment with nuance, multi-label categorization. Thinking helps with ambiguous cases.

🛡️ Customer Support

Handle complex support queries that need multi-step reasoning. Better quality than non-thinking budget models for tricky issues.

🔍 Research Synthesis

Synthesize information from multiple sources, compare arguments, identify gaps. Thinking capability enables deeper analysis at low cost.

When NOT to Use Gemini 3.5 Flash-Lite

Gemini 3.5 Flash-Lite is great for budget reasoning, but consider alternatives when:

Real-World Cost Scenario

Let's say you're building a document analysis pipeline that processes 25,000 documents per month, with an average of 3,000 input tokens and 800 output tokens (including thinking) per document:

ModelMonthly InputMonthly OutputTotal
Qwen 3.7 Flash$2.25$2.60$4.85
DeepSeek V4 Flash$10.50$5.60$16.10
Gemini 3.1 Flash-Lite$18.75$30.00$48.75
Gemini 3.5 Flash-Lite$22.50$50.00$72.50
Gemini 3 Flash$37.50$60.00$97.50
Claude Haiku 4.5$75.00$100.00$175.00

Gemini 3.5 Flash-Lite costs $72.50/month for 25K document analyses — 50% more than Gemini 3.1 Flash-Lite (no thinking) but 26% cheaper than Gemini 3 Flash (also no thinking). The thinking capability often produces better analysis, making the premium worthwhile.

Frequently Asked Questions

How much does Gemini 3.5 Flash-Lite cost?
$0.30 per million input tokens and $2.50 per million output tokens. Output pricing includes thinking tokens, so there are no surprise charges from hidden reasoning steps.
What is the context window of Gemini 3.5 Flash-Lite?
1M tokens. This is one of the largest context windows in the budget tier, matching other Google Flash models and exceeding GPT-5.4 nano's 400K context.
Does Gemini 3.5 Flash-Lite have thinking capability?
Yes. Gemini 3.5 Flash-Lite is Google's cheapest model with built-in thinking (reasoning) capability. Thinking tokens are included in the output price ($2.50/M), so you don't pay separately for reasoning steps.
How does it compare to Gemini 3.1 Flash-Lite?
Gemini 3.5 Flash-Lite ($0.30/$2.50) is slightly more expensive than Gemini 3.1 Flash-Lite ($0.25/$1.50) but adds thinking capability. The 3.5 version can reason through complex problems step-by-step, while 3.1 is a pure fast-inference model.
What can Gemini 3.5 Flash-Lite be used for?
Data analysis, code generation, document Q&A, nuanced classification, customer support, and research synthesis — any task that benefits from step-by-step reasoning at budget prices.

Calculate Your Gemini 3.5 Flash-Lite Costs

Enter your token counts and see exactly what you'd pay. Compare against 93 other models instantly.