Skip to main content
Back to News
Fish Audio, the AI voice startup behind open-source Fish Speech, has raised $52 million in seed fund...
Funding
2 min read

Fish Audio, the AI voice startup behind open-source Fish Speech, has raised $52 million in seed fund...

The AMW Read

Fish Audio is a new entrant in the AI voice segment; its large seed round and rapid ARR update the competitive landscape against ElevenLabs.
NoveltySignificance
Multimodal · Player Map

Fish Audio, the AI voice startup behind open-source Fish Speech, has raised $52 million in seed funding. The round was led by Coreline Ventures and Capital Today, with participation from 359 Capital, Play Time, HF0, 645 Ventures, and others. The company claims 8 million users and $21 million in ARR, and plans to expand into voice-native LLMs and real-time speech-to-speech translation.

Why it matters: Fish Audio's rapid ARR ramp to $21 million on seed funding exemplifies the 'fastest-ARR-ramp' pattern in AI voice, a segment where ElevenLabs recently raised $500M at an $11B valuation. The company's open-source origin (Fish Speech, 31K GitHub stars) and focus on emotional expression, 83-language support, and on-premises HIPAA deployment position it as a credible challenger in the enterprise voice AI market. This round also signals that AI voice generation is becoming a critical interface layer, with voice as the default modality for AI interaction.

Grounded take: Fish Audio hits the classic 'open-source to commercial' trajectory seen in foundation model startups, but with a twist: it targets the voice AI vertical, not general-purpose text. Its $52M seed is unusually large—more than most Series A rounds—and its ARR suggests strong product-market fit. The company's plan to make its flagship S2.1 Pro model free via API by end of August is a bold move to accelerate adoption and build a developer ecosystem, mirroring the 'hyperscaler-distribution moat' strategy. However, ElevenLabs remains the dominant player with deep pockets, and capital compression in AI audio could intensify. The open debate is whether voice AI will consolidate around a single platform or remain multi-model.

#AIvoice #TextToSpeech #SeedFunding #FishAudio #VoiceCloning #GenerativeAI #ElevenLabs

#Fish Audio#AI voice generation#seed funding#text-to-speech#voice cloning#ElevenLabs

How This Connects

Based on Multimodal · Player Map

  1. 4d agoMeshy reports $100M ARR as AI 3D assets enter production workflowsMeshy
  2. 6d agoStability AI targets music professionals with three audio models and editing softwareStability AI
  3. 1w agoWorld Labs agrees to $8.2 billion AMD acquisition as chipmaker expands into world modelsWorld Labs
  4. 0mo agoSuno replaces its music-generation lineup with Suno v6, trained on licensed catalogs from Warner Music Group, BMG, and Believe.Suno
  5. 0mo agoSuno Strikes Licensing Deals With Warner Music Group and BMG for New AI ModelsSuno
  6. 2mo agoFish Audio, the AI voice startup behind open-source Fish Speech, has raised $52 million in seed fund... · THIS ARTICLE

More news from Fish Audio

Stay updated with the latest news and announcements from Fish Audio.

View all Fish Audio news

Discover AI Startups

Explore 5,000+ AI companies with VC-grade analysis, funding data, and investment insights.

Explore Dashboard