Fish Audio has completed a $52 million seed round financing, unveiling the next-generation speech model S2.1 Pro.
July 29th, according to official sources, AI voice company Fish Audio announced the completion of a $52 million seed funding round, with specific investors yet to be disclosed, while officially launching the new generation of voice generation model S2.1 Pro.
Fish Audio stated that S2.1 Pro can clone voices using only a 5-second audio sample, with a generation speed about 2 times faster than Cartesia and a cost about one-sixth of ElevenLabs. Additionally, the model supports fine control of word-level emotions, intonation, speech rate, and more, being hailed as one of the most expressive voice models currently available.
Introduction indicates that many AI companies including HeyGen, LiveKit, Retell, Sanas, and OpenArt have already been using the Fish Audio model in production environments. The company stated that if an enterprise cannot reduce costs by at least 50% after adopting its speech AI service, they will provide one year of Fish Audio service for free.
To celebrate its first anniversary, Fish Audio also announced that it will offer users a free one-month trial of the S2.1 Pro model.