Gemini 3.5 Flash
Language AvailableCompare models
ComparingGemini 3.5 Flashwith
Pick a model above to see the comparison.
google-deepmind/gemini-3.5-flash · by Google DeepMind · decoder-only-transformer
Pricing — 1 offering(s)
Input tokens
- $1.50 / 1M tokens (input) Current 2026-07-07 → present
Output tokens
- $9.00 / 1M tokens (output) Current 2026-07-07 → present
Showing the active price and any recorded history. Full pricing history is available via the paid API — see API docs.
Capability profile
reasoning strong
coding strong
tool use strong
instruction following strong
context window strong
multilingual strong
speed strong
cost efficiency strong
Coding benchmarks
View full leaderboard →| Benchmark | Score | Harness | Source |
|---|---|---|---|
| swe-bench-verified ↗ | 78.8% 2026-05 | mini-swe-agent | 📊 leaderboard ↗ |
Operator guidance
The cost/latency option in the Gemini 3.x line, generally available (unlike 3.1 Pro's preview status). Step up to Gemini 3.1 Pro once it reaches GA for the hardest reasoning tasks; otherwise this is the default current-gen Gemini choice for most workloads.
Use cases
- High-throughput, low-cost tasks needing a large context window and frontier-adjacent reasoning
- Agentic and computer-use workflows (browser/desktop/mobile actions) at Flash-tier cost
Limitations
- Output pricing ($9.00) is notably higher than Gemini 2.5 Flash's ($2.50) for a Flash-tier model
- Capability ratings are qualitative, not a single cited benchmark run