ElevenLabs Launched New V4 Speech Models

The upgraded AI models expand language support to over 90 tongues and enable more precise expressive control.

Updated on Sept. 28, 2026 in Artificial Intelligence

Bold flat-color editorial illustration of a geometric sound-wave sculpture in navy, cream, and orange, representing synthetic voice technology.
ElevenLabs released its v4 and v4 Turbo speech models, featuring expanded expression controls and support for more than 90 languages. AI Illustration. Upload story photo >

Live Poll

Do you trust the rapid advancement of voice AI technology in your daily life?

ElevenLabs has released its v4 and v4 Turbo speech models, which feature expanded expression controls and require only 10 seconds of audio for voice cloning. These models, which now support more than 90 languages, are aimed at developers building voice agents and audio applications.

Why it matters

The company has scaled significantly, reaching an $11 billion valuation and an annualized revenue run rate exceeding $600 million. This release targets the growing demand for multilingual, expressive voice synthesis within enterprise voice-agent workflows.

The v4 models generate audio in real-time as the language model produces text, while allowing users to stack multiple inline expression tags to adjust delivery based on context.

The players

ElevenLabs

A developer of AI-powered speech synthesis and voice cloning software currently valued at $11 billion.

Sequoia

A venture capital firm that led the most recent $500 million funding round for ElevenLabs.

The details

The v4 system operates by integrating expression controls directly into the inference pipeline, allowing the model to adapt its delivery in sync with text generation. Users can influence prosody and tone by stacking tags that modulate the audio output as the model constructs the response. The platform also enables rapid voice cloning by processing just 10 seconds of source audio to capture a target persona.

Timeline

  1. September 28, 2026: ElevenLabs launched the v4 and v4 Turbo speech models.

The Tech Race

ElevenLabs is positioning its v4 models to capture a larger share of the voice agent market, where corporate adoption currently drives over 55% of its business. This release marks a departure from standard synthetic speech by prioritizing granular expression control alongside expanded global language support.

Developers can now utilize the new v4 models for projects requiring voice cloning with only 10 seconds of audio or multilingual support in over 90 languages. These capabilities are designed for integration into active voice agent workflows, though the specific availability of API tiers was not detailed.

The takeaway

ElevenLabs is moving to solidify its lead in the voice synthesis space by scaling its model complexity to meet enterprise requirements. Observers should track the company's expansion into India, Europe, and Brazil to see if these new language capabilities translate into further market dominance.

Further reading

For more on the latest trends in synthetic speech, visit Artificial Intelligence.

Live Poll

Do you trust the rapid advancement of voice AI technology in your daily life?