← All stories
● Covered by 1 source · 1 reportMedium impact1 neutral

ElevenLabs Launches v4 Speech Model with Enhanced Expression Control and 90-Language Support

🔄 Updated 3d ago
New to BrevFeed? We gather this story from every outlet covering it into one summary — ranked by real-world impact, not just the latest headline — so you never miss what matters. What is BrevFeed? →

Key points

  • ElevenLabs released v4 and v4 Turbo speech models.
  • New models support over 90 languages, up from 70.
  • Features include enhanced expression control and lower latency.
  • Voice cloning now requires only 10 seconds of audio.

New Speech Model Capabilities

ElevenLabs introduced its v4 and v4 Turbo speech models, building on the v3 model released last year. The new architecture allows for improved control over voice expression and faster voice cloning, requiring only 10 seconds of audio.

The v4 model handles voice identity more consistently over longer text passages and adjusts expressions based on text context. It expands on inline tags for expression definition, allowing users to stack multiple tags for sequential application.

Expanded Language Support

The new models increase language support from 70 to over 90 languages. ElevenLabs noted significant quality improvements for Japanese, Brazilian Portuguese, Mandarin, and Cantonese.

This expansion caters to a broader global user base and enhances the model's utility in diverse linguistic environments.

Enterprise Focus and Performance

ElevenLabs' enterprise calling business accounts for over 55% of its revenue. The v4 model is designed for voice agents, offering lower latency for more fluid conversations.

The model can generate audio as soon as the underlying Large Language Model (LLM) produces responses. It also manages confrontations, escalations, and holds differently to improve issue resolution in agent interactions.

Competitive Landscape and Company Growth

The speech model market is competitive, with companies like Cartesia, Deepgram, Fish Audio, Boson, and WellSaid Labs, alongside Google and OpenAI, developing expressive voice models.

ElevenLabs raised $500 million earlier this year, valuing the company at $11 billion. Its annualized revenue run rate has grown from approximately $330 million to over $600 million, and its headcount has exceeded 800 employees across various markets.

✨ This summary was generated by AI from the outlets' reporting listed below. It is not independently verified and may contain errors — check the original sources. How BrevFeed works →

The daily brief

One email each morning: the day's tech stories, clustered across outlets and summarized. No account needed.

One email a day. Unsubscribe in one click, any time.

Today's brief

Spend a few minutes, get the whole day. Every topic's top stories in one hands-free rundown — listen, watch, or read the transcript.

~34 min · 27 stories · Oct 02

▶ Play today's brief Listen on Spotify

New every morning, and the back catalogue is archived by date.

Reporting from

ElevenLabs released its v4 and v4 Turbo speech models, featuring improved expression control, lower latency, and support for over 90 languages. This update allows for more fluid voice agent interactions and better voice cloning, addressing growing demand in the enterprise sector.