ElevenLabs founders must be SHAKING right now...
for years they owned TTS because nothing open-source sounded human
that’s over... Fish Audio just launched S2.1 Pro and they’re one of the few voice AI companies that also has open-weight models
> clone any voice from a 15 second clip
> type [whispers] or [laughing nervously] into the script and it obeys
> switch languages mid conversation without switching models
the model you used to rent a subscription for is now has a public download model
tinker with the weights on your own gpu or use the api