Fish Audio Raises $52 Million Seed Funding to Scale Open-Source Voice AI Platform

Top Picks
Stay Ahead of the Market
Get the latest startup funding, hiring trends and global opportunities delivered to your inbox every week.
Fish Audio Secures $52 Million in Seed Funding
Palo Alto-based Fish Audio, a voice AI startup, has raised $52 million in Seed funding to expand its open-source voice AI platform and enterprise infrastructure.
The round was co-led by Coreline Ventures and another investor that was not identified in the provided source. Despite the size of the investment, the company continues to classify the round as Seed funding.
Founded approximately one year ago, Fish Audio has built a platform that combines open-weight voice models with enterprise-grade API services. The company says its platform now serves more than 8 million users and generates $21 million in annual recurring revenue (ARR).
An Open-Source Business Model for Voice AI
Fish Audio has adopted an open-source strategy by releasing several of its voice AI models publicly while monetizing enterprise infrastructure.
The company has released five voice models, including three open-source models, and offers its latest S2.1 Pro model free through its API until August 31.
According to the company, S2.1 Pro can:
Clone a voice from a five-second audio sample
Support 83 languages
Generate first audio in approximately 70 milliseconds
Revenue comes from enterprise customers that require low-latency performance, uptime guarantees, and commercial support.
Customers using Fish Audio's APIs include HeyGen, LiveKit, Retell, Sanas, and OpenArt.
Expanding Voice AI While Addressing Trust
A key asset for Fish Audio is its community-driven voice library, which contains more than two million user-submitted voices.
Users can contribute voices and receive compensation when those voices are used. However, the company has also faced concerns regarding unauthorized voice uploads.
According to the report, Fish Audio has introduced a dedicated dispute process and says voice removal requests are now completed in under three minutes after being reported.
The newly raised funding will support the continued expansion of the company's voice AI infrastructure as demand grows for enterprise voice applications.










