Fish Audio Closes $52 Million Seed Round, Launches S2.1 Pro Voice Model
AI voice startup Fish Audio has raised a $52 million seed funding round and publicly launched its S2.1 Pro model. The new system generates conversational audio in 90 milliseconds and can clone a voice using just five seconds of audio input.
The model supports 83 languages and includes word-level controls for adjusting emotion, pacing, and intonation. Fish Audio noted that several AI companies, including HeyGen and LiveKit, already run the technology in production. The company is also offering a promotional free trial to mark its first anniversary and has issued a cost-reduction guarantee for enterprise clients.
From the sources (3 posts)
@bdsqlszRT @FishAudio: Today we’ve raised $52M Seed and we are announcing the public launch of S2.1 Pro. >It can clone a voice from 5 seconds of…
@itsolelehmanni think voice AI is finally fast/expressive enough to make conversation simulators a massive new product category think use cases like: > an angry customer that trains support teams > a sales call you can replay with different answers > a
@fishaudioToday we’ve raised $52M Seed and we are announcing the public launch of S2.1 Pro. >It can clone a voice from 5 seconds of audio >2x faster than Cartesia & 1/6th the cost of Eleven Labs >most expressive model with word level control over e