Fish Audio publishes Fish Speech, an open-weight low-latency multilingual text-to-speech family with voice cloning, popular for self-hosted real-time speech.
Visit Fish AudioModels tracked
2
Open weight
2
API only
0
Avg score
77.8
Top benchmark
85.2
TTS Arena
Total HF downloads
4.2K
First release
Sep 2024
Latest release
Dec 2024
Ranked by composite score across benchmarks, popularity, efficiency, and versatility.
| # | Model | Modality | Score | Params | Released |
|---|---|---|---|---|---|
| 1 | audio | AA78.1 | — | Sep 2024 | |
| 2 |
The strongest model from this provider in each modality. Click a card to open the model page.
How this provider’s catalog splits across text, image, video, audio, and embedding.
When each model shipped, newest first.
Dec 4
Sep 10
Models with downloadable weights, ranked by composite score.
| # | Model | Modality | Score | Params | Released |
|---|---|---|---|---|---|
| 1 | audio | AA78.1 | — | Sep 2024 | |
| 2 |
Each family is a model series. Browse all releases in a family with their own leaderboard.
Fish Audio ships 2 models that we track. Of those, 2 are open weight and 0 are accessed through their API.
It depends on the model. See the open-weight section above for models you can run locally, and the API-only section for closed models priced per token.
By composite score, Fish Speech v1.4 is currently the top Fish Audio model in our directory. Composite score blends benchmarks, popularity, efficiency, and versatility.

Full Directory
Open the full directory to filter by hardware, capability, license, and benchmark score.

Or Browse by Family
See every release in a family side by side, with a timeline and a leaderboard. Useful when you have already picked a series and want the right size.
AA77.5 |
| — |
| Dec 2024 |
Composite grades across this provider’s catalog. Higher is better, blending benchmarks, popularity, and efficiency.
AA77.5 |
| — |
| Dec 2024 |