S1 is a multilingual text-to-speech model from Fish Audio. It is suited for voice applications that need broad emotional expression, using parenthetical controls to guide speaking style across its supported languages.
Modalities
Price
$15/M UTF-8 bytes
Released
Jul 29, 2026