Fish Audio Raises $52 Million Seed Funding to Scale Open-Source Voice AI Platform

Fish Audio Raises $52 Million Seed Funding to Scale Open-Source Voice AI Platform
GlobalNorth AmericaFunding
WorkNation
July 30, 2026

Fish Audio Secures $52 Million in Seed Funding

Palo Alto-based Fish Audio, a voice AI startup, has raised $52 million in Seed funding to expand its open-source voice AI platform and enterprise infrastructure.

The round was co-led by Coreline Ventures and another investor that was not identified in the provided source. Despite the size of the investment, the company continues to classify the round as Seed funding.

Founded approximately one year ago, Fish Audio has built a platform that combines open-weight voice models with enterprise-grade API services. The company says its platform now serves more than 8 million users and generates $21 million in annual recurring revenue (ARR).

An Open-Source Business Model for Voice AI

Fish Audio has adopted an open-source strategy by releasing several of its voice AI models publicly while monetizing enterprise infrastructure.

The company has released five voice models, including three open-source models, and offers its latest S2.1 Pro model free through its API until August 31.

According to the company, S2.1 Pro can:

  • Clone a voice from a five-second audio sample

  • Support 83 languages

  • Generate first audio in approximately 70 milliseconds

Revenue comes from enterprise customers that require low-latency performance, uptime guarantees, and commercial support.

Customers using Fish Audio's APIs include HeyGen, LiveKit, Retell, Sanas, and OpenArt.

Expanding Voice AI While Addressing Trust

A key asset for Fish Audio is its community-driven voice library, which contains more than two million user-submitted voices.

Users can contribute voices and receive compensation when those voices are used. However, the company has also faced concerns regarding unauthorized voice uploads.

According to the report, Fish Audio has introduced a dedicated dispute process and says voice removal requests are now completed in under three minutes after being reported.

The newly raised funding will support the continued expansion of the company's voice AI infrastructure as demand grows for enterprise voice applications.