Fish Audio, a voice AI startup developing expressive speech generation and audio models, raised $52 million in seed funding on its first anniversary. The round was led by new investors with participation from 359 Capital, Play Time, HF0, 645 Ventures, Parable, Carya Venture Partners, Alphaist Partners, and existing angel investors. The funding will accelerate research, developer tools, enterprise growth, and new audio AI capabilities.
Advancing Expressive Voice AI
Fish Audio focuses on natural voice generation that captures delivery and emotion rather than simply producing accurate speech. Over the past year, the company launched five AI models, reached $21 million in annual recurring revenue, attracted more than eight million users, and built a community library with over two million voice models supporting 83+ languages.
Expanding the Audio AI Stack
The company plans to expand beyond text-to-speech into speech-to-speech technology and Audio Language Models while strengthening APIs, enterprise products, and strategic partnerships. Enterprise customers use the platform for secure deployments, multilingual voice generation, and high-performance AI speech applications.
Supporting Developers and Enterprise
To mark the funding milestone, Fish Audio is offering free API access to its S2.1 Pro model through the end of August, alongside discounted creator plans. The company attributes the rollout to major improvements in inference efficiency, enabling lower operating costs and broader access to its voice AI technology.
