Ling-3.0-flash
Language Available Full comparison ↗ ant-group/ling-3-0-flash · by Ant Group
· mixture-of-experts
Pricing — 1 offering(s)
Input tokens
- $0.000 / 1M tokens (input) Historical 2026-07-24 → 2026-08-09
- $0.021 / 1M tokens (input) Current 2026-08-10 → present
Output tokens
- $0.000 / 1M tokens (output) Historical 2026-07-24 → 2026-08-09
- $0.063 / 1M tokens (output) Current 2026-08-10 → present
Showing the active price and any recorded history. Full pricing history is available via the paid API — see API docs.
See how Ling-3.0-flash fits into a cost-aware routing setup
See how →Capability profile
Operator guidance
Route here when cost and serving efficiency are the overriding constraints and the task is within reach of a ~5B-active model — Ant's efficiency play. For heavier reasoning, use the sibling Ring models or a larger flagship. All capability signal is vendor-reported (Ant's own announcement / inclusionAI benchmarks); there is no first-party Ant pricing console and no independent benchmark reproduction, so treat the "matches our 1T flagship" claim as directional.
Use cases
- Production-scale agent deployments where per-token cost dominates and a small active footprint is decisive
- High-throughput coding / tool-use pipelines at near-zero token cost
- Self-hosted deployment under MIT; pair with the Ring family for reasoning-heavy steps
Limitations
- All benchmark/capability claims are from Ant / inclusionAI's own materials — no independent reproduction found
- No verifiable first-party Ant Group pricing console; distributed purely through third-party hosts (OpenRouter, Novita, Vercel AI Gateway, ZenMux)
- 1M context is an ambition, not shipped — native window is 256K
- Registry pricing moved from a $0/$0 launch promo to a third-party steady-state rate via a price-monitor proposal; re-verify against a stable source