Introducing Muse Voice Transcribe, the first real-time audio perception model from Meta Superintelligence Labs.
Muse Voice Transcribe delivers real-time streaming ASR, diarization with 20+ speakers, and endpointing. It’s multilingual with seamless code-switching and improves accuracy with language, keyword, and context biasing.
The model ranks first on
@ArtificialAnlys streaming speech-to-text and on public diarization benchmarks.