Model ReleaseNew

Meta released Muse Voice Transcribe real-time speech transcription model

Meta's Superintelligence Labs has released Muse Voice Transcribe, a real-time speech transcription model designed to process audio in 80-millisecond chunks. According to benchmark findings from Artificial Analysis cited in the report, the model offers the highest streaming transcription accuracy at the lowest market price, intended to support continuous listening on wearable hardware such as camera glasses. Claims are as reported; this summary makes no determination about accuracy or significance.

First detected
Sep 6, 2026
Last updated
Sep 6, 2026

Moderate confidence

Based on a single independent report.

Limited corroboration

1 reporting source

What does this mean?

Corroboration measures how many genuinely independent sources support the event. Confidence measures how reliable the available evidence appears.

Newly detected

This development was detected recently and reporting may still arrive.

Follow this development to see meaningful updates as new evidence emerges.

Save keeps this for later. Follow tracks meaningful changes as new evidence emerges — it shapes your Following Feed, alerts, and digest eligibility, and doesn't promise an instant notification.

Why it matters

Low-latency, low-cost transcription with native speaker separation removes major technical and pricing barriers for developers building always-on conversational agents. However, integrating persistent real-time audio capture into consumer smart glasses will heighten privacy and ambient recording compliance concerns for enterprise and public environments.

Coverage

How this developed

  1. Sep 6, 2026

    1. Development detected

    2. New reporting added