← All models

Gemini 3.5 Flash

Language Available

google-deepmind/gemini-3.5-flash · by Google DeepMind · decoder-only-transformer

Pricing — 1 offering(s)

Input tokens

  • $1.50 / 1M tokens (input) Current 2026-07-07 → present

Output tokens

  • $9.00 / 1M tokens (output) Current 2026-07-07 → present

Showing the active price and any recorded history. Full pricing history is available via the paid API — see API docs.

Capability profile

reasoning strong
coding strong
tool use strong
instruction following strong
context window strong
multilingual strong
speed strong
cost efficiency strong

Coding benchmarks

View full leaderboard →
Benchmark Score Harness Source
swe-bench-verified ↗ 78.8% 2026-05 mini-swe-agent 📊 leaderboard ↗

Operator guidance

The cost/latency option in the Gemini 3.x line, generally available (unlike 3.1 Pro's preview status). Step up to Gemini 3.1 Pro once it reaches GA for the hardest reasoning tasks; otherwise this is the default current-gen Gemini choice for most workloads.

Use cases

Limitations

Citations