Meta Launches Muse Voice Transcribe for Real-Time Voice Dictation on Mac
The model supports 70+ languages and 20+ speakers, and Meta says it ranks first on a streaming speech-to-text leaderboard.
- On Tuesday, September 1, 2026, Meta launched Muse Voice Transcribe, a real-time audio perception model available via the Meta Model API for $3 per 1,000 audio-minutes, equivalent to $0.18 per hour.
- Meta Superintelligence Labs developed the system using "adaptive delay" technology, which adjusts listening time based on audio complexity. The model supports more than 70 languages with 25 validated at launch and handles 20-plus speakers.
- Developers can access the technology through Muse Code and the Meta Model API, while users experience the model within the Meta AI for Mac application. On Mac, holding the Fn key enables dictation into any application.
- Meta reports the model ranks first on the Artificial Analysis streaming speech-to-text leaderboard as of September 1. The release arrives less than a week after Google launched Google Gemini 3.5 Transcribe with similar capabilities.
- While Google integrates its model into Android, Meta has not yet detailed plans for similar flagship service integration. Meta CEO Mark Zuckerberg recently returned to X and shared examples of the model handling multiple speakers and languages simultaneously.
26 Articles
26 Articles
Meta Launches AI-Powered Muse Voice Transcribe With 5 Indian Language Support
Meta launches Muse Voice Transcribe, a real-time AI speech-to-text model that supports multilingual conversations and faster transcription. Meta’s new Muse Voice Transcribe uses AI to deliver real-time, multilingual speech transcription.
Meta just beat OpenAI and Google at real-time transcription
Meta’s Superintelligence Labs on Tuesday launched Muse Voice Transcribe, a new real-time speech recognition model that, at least on some benchmarks, outperforms virtually every other comparable model when it comes to working with speech in real time. Meta’s lab describes the model as its first “real-time audio perception model.” With Muse Spark, the company also recently shipped another speech-to-text capable model, though not one that specializ…
Meta launches Muse Voice Transcribe for real-time voice dictation on Mac
Meta is launching Muse Voice Transcribe, its first real-time audio perception model. It brings multilingual, streaming transcription to Meta AI for Mac, Muse Code, and developers through the Meta Model API. more…
Meta Releases Muse Voice Transcribe Speech Model – Recognizes over 70 Languages and over 20 Speakers
Meta has introduced a new AI model for real-time speech recognition and transcription, Muse Voice Transcribe. Developed by Meta Superintelligence Labs, the model is part of the Muse Spark model family and is already running ... Read more
Coverage Details
Bias Distribution
- 60% of the sources lean Left
Factuality
To view factuality data please Upgrade to Premium












