ElevenLabs is a Foundation Models - Voice & Music company founded in 2022 and based in London, United Kingdom. It has raised $801M in total funding, most recently a Series D in 2026 at a $11B valuation.
| Date | Stage | Amount | Valuation | Lead investors |
|---|---|---|---|---|
| Mar 12, 2026 | Undisclosed | $20M | — | Robinhood |
| Feb 4, 2026 | Series D | $500M | $11B |
| Sequoia Capital |
| Jan 30, 2025 | Series C | $180M | $3.3B | Andreessen Horowitz, ICONIQ Growth |
| Jan 22, 2024 | Series B | $80M | $1.1B | Andreessen Horowitz, Sequoia Capital |
| Jun 1, 2023 | Series A | $19M | $100M | Andreessen Horowitz |
| Jan 23, 2023 | Pre-Seed | $2M | — | Credo Ventures |
The most expressive text-to-speech model supporting 29 languages. Captures emotional nuance, context awareness, and natural delivery through advanced AI algorithms that understand vocal emotion, intonation, and pacing. Designed for long-form content including audiobooks, film, dramatic voiceovers, and character-driven narratives. Enables creators and enterprises to generate lifelike speech that synthesizes emotion and intent, with support for inline audio tags like [whispers], [laughs], and [excited] for fine-grained control over delivery.
AI music generation platform launched August 2025 that enables users to generate studio-grade, commercially-safe music from natural language prompts. Features proprietary AI trained on licensed data from Merlin Network and Kobalt Music Group, supporting multiple genres, styles, and languages including English, Spanish, German, and Japanese. Allows creation of complete tracks with or without vocals, and includes editing capabilities for individual song sections. Cleared for commercial use across film, television, podcasts, advertising, gaming, and social media.
Developer platform for deploying conversational voice agents at scale, launched November 2024. Enables businesses to build interactive voice experiences that listen, read, and respond naturally across phone, chat, email, and WhatsApp. Features low-latency response times using Flash v2.5 model achieving sub-500ms end-to-end latency. Includes testing, monitoring, compliance tools, and integrations necessary for enterprise deployment. Empowers companies to deliver seamless customer experiences through voice-first interfaces for support, sales, and marketing applications.
All-in-one AI-native creative workspace combining voice, music, sound effects, image, and video generation. Empowers creators and marketers to generate and edit speech, music, images, and video across 70+ languages. Provides access to 10,000+ human-like AI voices plus licensing of iconic voices. Includes features like voice cloning, sound effect generation, dubbing, and transcription. Designed for professional and enterprise use with APIs, SDKs, and iOS/Android apps enabling seamless integration into production workflows and end-to-end content creation pipelines.