Twitter/X

Fish Audio announced on 2026-07-29 it raised $52M Seed and publicly launched S2.1…

Brief

Fish Audio announced on 2026-07-29 it raised a $52M seed round and publicly launched S2.1 Pro, which it says can clone a voice from 5 seconds of audio. The company claims 2× speed vs. Cartesia, 1/6th the cost of Eleven Labs, word-level expressive control, production use at HeyGen/LiveKit/Retell/Sanas/OpenArt, and several cost-reduction promos.

Why it matters

Fish Audio announced on 2026-07-29 it raised $52M Seed and publicly launched S2.1 Pro, claiming the model can clone a voice from 5 seconds of audio.

Key details

  • Company claims S2.1 Pro is 2× faster than Cartesia, costs 1/6th of Eleven Labs, provides word-level control over emotion, intonation, and pacing, and runs in production at HeyGen, LiveKit, Retell, Sanas, and OpenArt.
  • Commercial offers include a guarantee to cut a business's voice-AI costs by 50% or provide one year free, a demo link (s.fish.audio/tmapke), and a one-month free S2.1 Pro promotion for Fish’s first birthday for likes/retweets/comments of “Fish.”
Source evidence

Check out @FishAudio's official announcement:

Fish Audio (@FishAudio)

Today we’ve raised $52M Seed and we are announcing the public launch of S2.1 Pro.

>It can clone a voice from 5 seconds of audio
>2x faster than Cartesia & 1/6th the cost of Eleven Labs
>most expressive model with word level control over emotion, intonation, pacing etc

We support frontier AI companies including HeyGen, LiveKit, Retell, Sanas, and OpenArt all run our model in production.

If you're a business and we can't cut your voice AI costs by 50%, we'll give you 1 year of Fish Audio for free.
Book a demo: s.fish.audio/tmapke

To celebrate our first birthday, we'll give you 1 month of S2.1 Pro for free. Like, retweet, and comment “Fish” to get it.

Video

— https://nitter.net/FishAudio/status/2082152596739862853#m