Fish Audio
Open-weight multilingual TTS and voice cloning.
Full Fish Audio record →Head to head
Fish Audio (Fish Audio) and Whisper (OpenAI) both sit in Audio. Whisper shipped more tracked changes in the last 30 days (2 vs 0). Every value below comes from the latest crawl of the vendors' own pages.
| Attribute | Fish Audio Fish Audio | Whisper OpenAI |
|---|---|---|
| Segment | Audio | Audio |
| Version | S1 | gpt-transcribe |
| Entry cost | Open weights; API from $15/M chars | Pricing not specified in the document |
| Pricing tiers | — | Open Source Model free |
| Model stack | Fish Speech S1 | gpt-transcribe · Whisper · Whisper large-v3 · whisper-1 |
| Context window | — | 30 second windows |
| Public API | ||
| Routes models | ||
| Changes / 30d | 0 | 2 |
| Origin | China | Global |
| Capabilities | Voice cloning, Multilingual, Self-host | 99 languages, Translation to English, Word timestamps, Open weights, Robust to noise, File transcription, Language detection, Support for mp3, mp4, mpeg, mpga, m4a, wav, and webm formats |
Open-weight multilingual TTS and voice cloning.
Full Fish Audio record →A widely deployed multilingual transcription and translation model.
Full Whisper record →