Fish Audio S1
Audio Available fish-audio/s1 · by Fish Audio · neural-tts
Pricing — 1 offering(s)
Text characters
- $15.00 / 1M characters Current 2026-07-30 → present
Showing the active price and any recorded history. Full pricing history is available via the paid API — see API docs.
Capability profile
voice naturalness strong
language support moderate
voice variety strong
streaming latency unknown
cloning support strong
Benchmarks
| Benchmark | Score | Config | Source |
|---|---|---|---|
| Word error rate (vendor-reported) | 0.8 % | Fish Audio's own reported WER. Not independently re-measured by Modelglass. | source ↗ |
| TTS-Arena2 ranking (independent, community leaderboard) | — | Ranked #1 at time of writing per Fish Audio's own docs — an independent/community leaderboard, not a vendor-run benchmark, though the ranking snapshot itself wasn't independently re-verified by Modelglass. | — |
Operator guidance
Fish Audio's own docs label S1 "Previous Model" and recommend s2.1-pro for all new projects — S1 is kept available only "for existing integrations," priced identically to S2 Pro/S2.1 Pro. Use S2.1 Pro instead for anything new; S1 is documented here for operators already integrated against it.
Use cases
- Existing integrations already built on S1 (Fish Audio's own guidance — new projects should use S2.1 Pro instead)
- Cost-sensitive TTS with strong naturalness at S2/S2.1 Pro's same price point
- Expressive narration via 64+ emotion/tone markers using (parenthesis) syntax
Limitations
- Fish Audio itself recommends against using S1 for new projects — kept only for existing integrations
- Narrower language support (13) than S2 Pro/S2.1 Pro (80+/83) or ElevenLabs' models
- No published time-to-first-audio guarantee, unlike S2 Pro/S2.1 Pro's 100ms SLA
- Fish Audio's native billing unit is UTF-8 bytes, not characters — see registry pricing notes
Citations
Compare models
ComparingFish Audio S1with
Pick a model above to see the comparison.