模型

ElevenLabs发布v4语音模型

ElevenLabs' new v4 speech model makes AI voices more expressive and consistent

精选理由

ElevenLabs新v4语音模型,笑语音准,长音频一致,Turbo响应快150ms,已超Gemini。

ElevenLabs推出v4语音模型,能更准确捕捉笑声和低语指令,保持长音频一致性。其Turbo变体响应速度达150毫秒,专为实时语音助手设计。在Voice Arena排行榜上,v4超越Cartesia和Google Gemini。

原文 · Decoder

ElevenLabs' new v4 speech model makes AI voices more expressive and consistent

Elevenlabs' new speech model, Eleven v4, follows cues for laughter and whispering more accurately and keeps voices consistent across long productions like audiobooks. Its Turbo variant starts speaking in 150 milliseconds and is built for real-time voice agents. On Artificial Analysis' Voice Arena leaderboard, v4 ranks ahead of Cartesia and Google's Gemini. The article ElevenLabs' new v4 speech model makes AI voices more expressive and consistent appeared first on The Decoder .