← All stories
● Covered by 1 source · 1 reportMedium impact1 neutral

Fish Audio Raises $52M Seed for AI Voice Model Development

🔄 Updated 1d ago
New to BrevFeed? We gather this story from every outlet covering it into one summary — ranked by real-world impact, not just the latest headline — so you never miss what matters. What is BrevFeed? →

Key points

  • Fish Audio raised $52 million in seed funding.
  • The company develops AI voice models for creative and enterprise use cases.
  • It has over 8 million users and $21 million in annual recurring revenue.
  • Fish Audio offers both open-source and paid API models.

Funding Details

Fish Audio announced it has raised $52 million in a seed funding round. The round was led by Coreline Ventures and Capital Today, with participation from 359 Capital, Parable, Play Time, Alphalist Partners, Bayhouse Ventures, Carya Venture Partners, and HF0.

AI Voice Model Offerings

Fish Audio develops AI voice models designed for both creative applications requiring expressiveness and enterprise uses needing steerability for customer support and sales operations. The company's library includes over 15,000 natural language controls for its voice models.

Growth and Adoption

Since its launch last year, Fish Audio has accumulated over 8 million users across its open-source and hosted model versions. The company reports an annual recurring revenue of $21 million. It has released five models in the past year, including four speech generation models and one speech-to-text model. Three of its speech generation models are open-source, while the S2.1 Pro model is available via a paid API.

Enterprise and Creator Solutions

Fish Audio offers paid monthly plans for creators and teams, which include generation minutes and voice cloning features. An enterprise version of its APIs and platform is also available, with companies like HeyGen and Sanas utilizing its services. The company notes that different enterprises have varying needs, such as realism for AI avatars, expressiveness for gaming characters, or natural-sounding, low-latency voices for calls.

Origin and Challenges

The company originated from a project by former NVIDIA researcher Shijia Liao, who developed a voice generation model due to dissatisfaction with existing synthetic voices. This model was open-sourced, leading to the Fish Speech repository on GitHub, which has over 31,000 stars. Fish Audio previously faced issues regarding user-submitted voices, with some creators alleging their voices were used without consent, which the company addressed through a DMCA takedown process.

✨ This summary was generated by AI from the outlets' reporting listed below. It is not independently verified and may contain errors — check the original sources. How BrevFeed works →

The daily brief

One email each morning: the day's tech stories, clustered across outlets and summarized. No account needed.

One email a day. Unsubscribe in one click, any time.

Today's brief

Spend a few minutes, get the whole day. Every topic's top stories in one hands-free rundown — listen, watch, or read the transcript.

~7 min · 6 stories · Aug 15

▶ Play today's brief Listen on Spotify

New every morning, and the back catalogue is archived by date.

Reporting from

Fish Audio, a startup specializing in AI-generated voice models, secured $52 million in seed funding to expand its offerings for creators and enterprises. The company provides expressive and steerable AI voice models, with over 8 million users and $21 million in annual recurring revenue.