Skip to content
The Decoder· Jonathan Kemper·· 2 d ago

ElevenLabs' new v4 speech model makes AI voices more expressive and consistent

ElevenLabs' new v4 speech model makes AI voices more expressive and consistent

AI summary

ElevenLabs released Eleven v4, which follows emotion, pause and sound-effect tags in scripts more accurately and maintains a consistent voice in long productions. The architecture also powers Turbo, starting speech output in around 150 milliseconds in official tests, compared with 262 milliseconds for Cartesia Sonic 3.6 and 814 milliseconds for OpenAI GPT-4o mini TTS.

Selection record

Not admittedSum of both 149 < twice the threshold 152

Source tier
Media and individuals; this tier's threshold is 76
Pre-filter
passed:ElevenLabs发布AI语音模型v4

A model scores each item twice, independently, against one written standard, out of 100. An item is admitted only when the two scores add up to twice the threshold. The threshold is set per source tier.

Source: The Decoder · the-decoder.com