ElevenLabs' new v4 speech model makes AI voices more expressive and consistent
Elevenlabs' new speech model, Eleven v4, follows cues for laughter and whispering more accurately and keeps voices consistent across long productions like audiobooks. Its Turbo variant starts speaking in 150 milliseconds and is built for real-time voice agents. On Artificial Analysis' Voice Arena l…
Read the full story at The Decoder ↗
Timeline · 1 report
- 2026-09-29 14:45 · The Decoder
ElevenLabs' new v4 speech model makes AI voices more expressive and consistent