← All models

Gemini 2.5 Flash

Language Available Full comparison ↗

google-deepmind/gemini-2.5-flash · by Google DeepMind · decoder-only-transformer

Pricing — 1 offering(s)

Input tokens

  • $0.30 / 1M tokens (input) Current 2026-06-12 → present

Output tokens

  • $2.50 / 1M tokens (output) Current 2026-06-12 → present

Showing the active price and any recorded history. Full pricing history is available via the paid API — see API docs.

See how Gemini 2.5 Flash fits into a cost-aware routing setup

See how →

Capability profile

reasoning moderate
coding moderate
tool use strong
instruction following strong
context window strong
multilingual strong
speed strong
cost efficiency strong

Coding benchmarks

View full leaderboard →
Benchmark Score Harness Source
swe-bench-verified ↗ 54.0% 2025-09 agentless 🏢 vendor ↗
aider-polyglot ↗ 55.1% 2025-05 aider 📊 leaderboard ↗

Operator guidance

The cost/latency option in the Gemini 2.5 line with a 1M-token window. Step up to Gemini 2.5 Pro for the hardest reasoning/coding.

Use cases

Limitations

Citations