congrats to the whole fish audio team. ive collaborated with their engineers extensively and want to call out something beyond the model: the professionalism and engineering discipline of this team is genuinely exceptional. the models speak for themselves - expressive TTS with word-level control at 1/6th the cost is a real achievement - but its the people behind them that impress me most. from open source to production-critical infrastructure for HeyGen, LiveKit, and others in a single year. thank you for being such a great partner @FishAudio
Fish Audio (@FishAudio)
Today we’ve raised $52M Seed and we are announcing the public launch of S2.1 Pro.
>It can clone a voice from 5 seconds of audio
>2x faster than Cartesia & 1/6th the cost of Eleven Labs
>most expressive model with word level control over emotion, intonation, pacing etc
We support frontier AI companies including HeyGen, LiveKit, Retell, Sanas, and OpenArt all run our model in production.
If you're a business and we can't cut your voice AI costs by 50%, we'll give you 1 year of Fish Audio for free.
Book a demo: s.fish.audio/tmapke
To celebrate our first birthday, we'll give you 1 month of S2.1 Pro for free. Like, retweet, and comment “Fish” to get it.
Video
— https://nitter.net/FishAudio/status/2082152596739862853#m