ElevenLabs' new v4 speech model makes AI voices more expressive and consistent

Elevenlabs' new speech model, Eleven v4, follows cues for laughter and whispering more accurately and keeps voices consistent across long productions like audiobooks. Its Turbo variant starts speaking in 150 milliseconds and is built for real-time voice agents. On Artificial Analysis' Voice Arena leaderboard, v4 ranks ahead of Cartesia and Google's Gemini. The article ElevenLabs' new v4 speech…
This is a summary curated by AIFuture. Read the complete article at the original source:
Read the full story on The Decoder