OPEN SOURCE KEEPS WINNING
Fish Audio just dropped S2.1 Pro, and it's built specifically for multi-agent voice infrastructure.
Look at these specs:
→ 2x faster than Cartesia
→ 1/6th the cost of ElevenLabs
→ 83+ languages
→ word level control over emotion, pacing and delivery
→ on prem with zero data retention
If you build voice agents, you have to check it out
Fish Audio (@FishAudio)
Today we’ve raised $52M Seed and we are announcing the public launch of S2.1 Pro.
>It can clone a voice from 5 seconds of audio
>2x faster than Cartesia & 1/6th the cost of Eleven Labs
>most expressive model with word level control over emotion, intonation, pacing etc
We support frontier AI companies including HeyGen, LiveKit, Retell, Sanas, and OpenArt all run our model in production.
If you're a business and we can't cut your voice AI costs by 50%, we'll give you 1 year of Fish Audio for free.
Book a demo: s.fish.audio/tmapke
To celebrate our first birthday, we'll give you 1 month of S2.1 Pro for free. Like, retweet, and comment “Fish” to get it.
Video
— https://nitter.net/FishAudio/status/2082152596739862853#m