Signal

ElevenLabs launches v4 and Turbo audio-model generation

Eleven v4 and v4 Turbo introduce a new speech-model architecture and broader expressive control. The September 28 release was updated October 1; advertised latency is provider-measured.

1 min read
ElevenLabs' dated launch page for Eleven v4 and v4 Turbo
ElevenLabs announces its Eleven v4 audio-model generation · Credit: ElevenLabs View source

ElevenLabs released Eleven v4 and its low-latency v4 Turbo variant on September 28, with the announcement updated October 1. The company describes an entirely new text-to-speech architecture, stronger emotional direction and support for more than 90 languages.

Availability and latency are separate claims

The release says the models are available through ElevenCreative, ElevenAgents and ElevenAPI. ElevenLabs reports roughly 100 milliseconds median inference latency and 150 milliseconds time to first speech for Turbo. Its comparison subtracts network latency, so this is not an end-to-end user latency guarantee.

The major audio generation is distinct from the company's employee tender. Expressiveness, cloning consistency and comparative speed remain provider claims; this review did not independently reproduce them.

Sources