ElevenLabs' new v4 speech model makes AI voices more expressive and consistent ElevenLabs released Eleven v4, a speech model that follows cues for laughter and whispering more accurately and maintains voice consistency across long productions such as audiobooks. The Turbo variant begins speaking in 150 milliseconds and targets real-time voice agents, and v4 ranks ahead of Cartesia and Google's Gemini on Artificial Analysis' Voice Arena leaderboard. Elevenlabs' new speech model, Eleven v4, follows cues for laughter and whispering more accurately and keeps voices consistent across long productions like audiobooks. Its Turbo variant starts speaking in 150 milliseconds and is built for real-time voice agents. On Artificial Analysis' Voice Arena leaderboard, v4 ranks ahead of Cartesia and Google's Gemini. The article ElevenLabs' new v4 speech model makes AI voices more expressive and consistent https://the-decoder.com/elevenlabs-new-v4-speech-model-makes-ai-voices-more-expressive-and-consistent/ appeared first on The Decoder https://the-decoder.com .