Mark RadarMARK RADAR
About
EN
Sign in
Event File AI Meta

Meta Launches Muse Voice Transcribe for Real-Time Speech Recognition

3 reports · First detected 2026-09-02 · Last active 2026-09-02

Meta Superintelligence Labs has introduced Muse Voice Transcribe, a real-time audio perception model that combines streaming automatic speech recognition, speaker diarization and endpoint detection in a single system. The unified design is intended to reduce the complexity and latency involved in linking separate speech tools, making it relevant for voice assistants, live captions and meeting transcription. The release also underscores Meta’s push to build multilingual audio infrastructure alongside its broader artificial-intelligence efforts.

Muse Voice Transcribe supports more than 70 languages and can distinguish over 20 speakers in multi-person audio, according to the release. Meta has made the model available through the Meta Model API and on Mac computers. On macOS, users can hold the Fn key to activate system-wide voice input, extending real-time transcription across applications and text fields without requiring each app to integrate a separate speech-recognition service.

All Coverage

3 original reports

The Backstory

The history behind this event

No historical echoes for this signal

Mark Radar|MARK RADAR

If you search news on Google, you can set Mark Radar as a preferred source—our coverage will show up more often in your results. Set as preferred source on Google →

All times are in Taipei time (GMT+8)