Gemini 2.0 Flash
Language Available Full comparison ↗ google/gemini-2.0-flash · by Google DeepMind
· decoder-only-transformer
Pricing — 1 offering(s)
Input tokens
- $0.10 / 1M tokens (input) Current 2025-02-05 → present
Output tokens
- $0.40 / 1M tokens (output) Current 2025-02-05 → present
Showing the active price and any recorded history. Full pricing history is available via the paid API — see API docs.
See how Gemini 2.0 Flash fits into a cost-aware routing setup
See how →Capability profile
reasoning moderate
coding moderate
tool use strong
instruction following strong
context window strong
multilingual strong
speed strong
cost efficiency strong
Operator guidance
Budget workhorse for Google's API. 10x cheaper on input than Gemini 2.5 Flash. Route here for high-volume workloads where Gemini 2.5 quality is not required. 1M context is a strong differentiator for long-doc tasks.
Use cases
- High-volume, cost-sensitive text generation
- Long-document processing (1M token context)
- Agentic pipelines needing fast, cheap completions
Limitations
- Previous generation; Gemini 2.5 Flash is meaningfully stronger on reasoning
- Capability ratings are qualitative, not a single cited benchmark run
Citations
Compare models
ComparingGemini 2.0 Flashwith
Pick a model above to see the comparison.