According to official announcement, AI voice company Fish Audio raised $52 million in seed funding and launched its next-generation voice model S2.1 Pro on July 29. The model can clone voices from just five-second audio samples, with generation speed approximately 2x faster than Cartesia and production costs about one-sixth of ElevenLabs.
S2.1 Pro supports word-level emotion, tone, and speed control and is described as one of the most expressive voice models available. Companies including HeyGen, LiveKit, Retell, Sanas, and OpenArt have already adopted Fish Audio models in production environments. To mark its one-year anniversary, Fish Audio is offering free one-month access to S2.1 Pro.