MODELS
by Fish Audio
Expressive text-to-speech model with natural-language emotion and prosody control, available through Fish Audio app/API and an open-source release.
Fish Audio S2 Pro is an expressive text-to-speech model from Fish Audio. The supplied Fish Audio sources describe it as supporting natural-language control over emotion and prosody, and as being available through the Fish Audio app/API as well as an open-source release. Official discovery sources include the Fish Audio launch/blog page, Fish Audio documentation and quickstart pages, and the fishaudio/fish-speech GitHub repository.
Use when you need a text-to-speech model with natural-language emotion and prosody control. Use when evaluating Fish Audio’s app/API offering alongside its open-source Fish Speech repository. Use when building or comparing catalog entries for expressive speech synthesis models.
Last updated Jul 31, 2026