Skip to main content
Model ReleaseNew

ElevenLabs released Eleven v4 speech model

ElevenLabs has released Eleven v4, an updated speech model that improves voice consistency over long productions and responds more accurately to emotional cues like whispering and laughter. The release includes a Turbo variant featuring a 150-millisecond latency for real-time voice agents. Claims are as reported; this summary makes no determination about accuracy or significance.

First detected
Sep 29, 2026
Last updated
Sep 29, 2026

Newly detected

This development was detected recently and reporting may still arrive.

Follow this development to see meaningful updates as new evidence emerges.

Save keeps this for later. Follow tracks meaningful changes as new evidence emerges — it shapes your Following Feed, alerts, and digest eligibility, and doesn't promise an instant notification.

Why it matters

Sub-200 millisecond response times and emotional nuance are essential benchmarks for deploying interactive voice agents and scalable audio production. The updated rankings indicate accelerating competition among specialized audio AI developers and frontier model providers in speech synthesis performance.

Coverage

Primary/vendor sources vs independent reporting

Primary / vendor source: information published directly by the company, organization, government body or project involved. Useful as a primary source, but not independent confirmation.

Independent reporting: reporting or analysis from a source independent of the organization making the underlying claim.

How this developed

  1. Sep 29, 2026

    1. Development detected

    2. New reporting added