← All models

Olmo 3 32B Think

Language Available Full comparison ↗

allenai/olmo-3-32b-think · by Allen Institute for AI (Ai2) · decoder-only-transformer

Pricing — 1 offering(s)

Input tokens

  • $0.15 / 1M tokens (input) Current 2025-11-20 → present

Output tokens

  • $0.50 / 1M tokens (output) Current 2025-11-20 → present

Showing the active price and any recorded history. Full pricing history is available via the paid API — see API docs.

See how Olmo 3 32B Think fits into a cost-aware routing setup

See how →

Capability profile

reasoning moderate
instruction following moderate
coding moderate
tool use moderate
context window weak
multilingual weak
speed moderate
cost efficiency strong

Operator guidance

Choose Olmo 3 32B Think when full openness (data + recipe + checkpoints) or self-hosting is the requirement and mid-tier reasoning accuracy is acceptable. For maximum reasoning quality at similar size, Qwen 3 32B and proprietary reasoning models score higher. If you want Ai2's openness with the strongest reasoning, look at the newer Olmo 3.1 32B Think instead of this snapshot. Not the choice for long-context, multilingual, or latency-critical work.

Use cases

Limitations

Citations