The Decoder· Jonathan Kemper·· 2 d ago
ElevenLabs' new v4 speech model makes AI voices more expressive and consistent
ElevenLabs' new v4 speech model makes AI voices more expressive and consistent
AI summary
ElevenLabs released Eleven v4, which follows emotion, pause and sound-effect tags in scripts more accurately and maintains a consistent voice in long productions. The architecture also powers Turbo, starting speech output in around 150 milliseconds in official tests, compared with 262 milliseconds for Cartesia Sonic 3.6 and 814 milliseconds for OpenAI GPT-4o mini TTS.
Selection record
Threshold 76Media and individualsFirst 78Second 71
Not admittedSum of both 149 < twice the threshold 152
- Source tier
- Media and individuals; this tier's threshold is 76
- Pre-filter
- passed:ElevenLabs发布AI语音模型v4
A model scores each item twice, independently, against one written standard, out of 100. An item is admitted only when the two scores add up to twice the threshold. The threshold is set per source tier.
Source: The Decoder · the-decoder.com