ElevenLabs introduced its v4 and v4 Turbo speech models, building on the v3 model released last year. The new architecture allows for improved control over voice expression and faster voice cloning, requiring only 10 seconds of audio.
The v4 model handles voice identity more consistently over longer text passages and adjusts expressions based on text context. It expands on inline tags for expression definition, allowing users to stack multiple tags for sequential application.
The new models increase language support from 70 to over 90 languages. ElevenLabs noted significant quality improvements for Japanese, Brazilian Portuguese, Mandarin, and Cantonese.
This expansion caters to a broader global user base and enhances the model's utility in diverse linguistic environments.
ElevenLabs' enterprise calling business accounts for over 55% of its revenue. The v4 model is designed for voice agents, offering lower latency for more fluid conversations.
The model can generate audio as soon as the underlying Large Language Model (LLM) produces responses. It also manages confrontations, escalations, and holds differently to improve issue resolution in agent interactions.
The speech model market is competitive, with companies like Cartesia, Deepgram, Fish Audio, Boson, and WellSaid Labs, alongside Google and OpenAI, developing expressive voice models.
ElevenLabs raised $500 million earlier this year, valuing the company at $11 billion. Its annualized revenue run rate has grown from approximately $330 million to over $600 million, and its headcount has exceeded 800 employees across various markets.
✨ This summary was generated by AI from the outlets' reporting listed below. It is not independently verified and may contain errors — check the original sources. How BrevFeed works →
One email each morning: the day's tech stories, clustered across outlets and summarized. No account needed.
One email a day. Unsubscribe in one click, any time.
Spend a few minutes, get the whole day. Every topic's top stories in one hands-free rundown — listen, watch, or read the transcript.
▶ Play today's briefNew every morning, and the back catalogue is archived by date.
ElevenLabs released its v4 and v4 Turbo speech models, featuring improved expression control, lower latency, and support for over 90 languages. This update allows for more fluid voice agent interactions and better voice cloning, addressing growing demand in the enterprise sector.