Eleven v4: Our most expressive text-to-speech AI model yet

The text-to-speech models support over 90 languages, inline emotion prompting, and a median time to first speech of ~150ms on the Turbo variant.

Developers and creators get more expressive multi-speaker voice synthesis with lower latency and fine-grained natural language delivery controls.

Eleven v4: Our most expressive text-to-speech AI model yet

Sources

Read this as text

Back to the AI news