RELEASE · MODELS · #1202
ElevenLabs releases Eleven v4 speech model with more expressive, consistent voices
ElevenLabs released Eleven v4, a speech model that better follows direction cues (laughter, whispers, sound effects), maintains voice consistency across long productions, supports over 90 languages and up to 10,000 characters per request, and restores professional voice cloning; an optimized low-latency 'Turbo' variant begins producing audio in about 150 ms. Eleven v4 is available via ElevenAgents, ElevenCreative and the API, with published benchmark and pricing comparisons to rivals and regionally configurable data storage options for enterprise customers.
KEY POINTS
- ElevenLabs released Eleven v4, a speech model that better follows direction cues (laughter, whispers, sound effects), maintains voice consistency across long productions, supports over 90 languages and up to 10,000 characters per request, and restores professional voice cloning; an optimized low-latency 'Turbo' variant begins producing audio in about 150 ms.
- Eleven v4 is available via ElevenAgents, ElevenCreative and the API, with published benchmark and pricing comparisons to rivals and regionally configurable data storage options for enterprise customers.
- v4's improved expressivity, cross‑language cloning durability, and a low‑latency Turbo option materially boost capabilities for dubbing, real‑time voice agents, and commercial TTS use cases.
WHY IT MATTERS
v4's improved expressivity, cross‑language cloning durability, and a low‑latency Turbo option materially boost capabilities for dubbing, real‑time voice agents, and commercial TTS use cases.