← All models

Llama 4 Maverick

Language Available Full comparison ↗

meta/llama-4-maverick · by Meta · mixture-of-experts

Pricing — 1 offering(s)

Input tokens

  • $0.15 / 1M tokens (input) Historical 2025-04-05 → 2026-07-21
  • $0.27 / 1M tokens (input) Current 2026-07-22 → present

Output tokens

  • $0.60 / 1M tokens (output) Historical 2025-04-05 → 2026-07-21
  • $0.85 / 1M tokens (output) Current 2026-07-22 → present

Showing the active price and any recorded history. Full pricing history is available via the paid API — see API docs.

See how Llama 4 Maverick fits into a cost-aware routing setup

See how →

Capability profile

reasoning moderate
coding moderate
tool use moderate
instruction following strong
context window strong
multilingual moderate
speed strong
cost efficiency strong

Coding benchmarks

View full leaderboard →
Benchmark Score Harness Source
swe-bench-verified ↗ 21.0% 2025-07 mini-swe-agent 0.0.0 📊 leaderboard ↗
livecodebench ↗ 43.4% 2025-04 📊 leaderboard ↗
aider-polyglot ↗ 15.6% 2025-04 aider 📊 leaderboard ↗
swe-bench-pro 5.2% 2026-01 swe-agent 📊 leaderboard ↗
terminal-bench-2-1 7.9% 2026-08 terminus-2 🔬 independent ↗
bigcodebench ↗ 28.4% 2025-04 📊 leaderboard ↗

Operator guidance

Route here for the longest context available at this price point, or for vision tasks where GPT-4o cost is prohibitive. At the same $0.15/$0.60 price as GPT-4o mini, Maverick's open weights and 1M-token context are the differentiators. Step up to frontier closed models for the hardest reasoning tasks.

Use cases

Limitations

Citations