Tech News
← Home  ·  All topics

Muse Voice Transcribe

3 GoKawiil briefs on this topic

Meta launches Muse Voice Transcribe at $0.18 per audio hour with 20+ speaker diarization

Meta Superintelligence Labs has released Muse Voice Transcribe, a real-time speech-to-text model that streams transcription, detects sentence endpoints, and identifies more than 20 distinct speakers without a separate processing step. Trained on over 70 languages (25 thoroughly validated), the model handles long recordings beyond an hour and supports code-switching between languages, all offered via API at $0.18 per hour of audio.

Meta launches Muse Voice Transcribe, a real-time multilingual speech-to-text model

Meta unveiled Muse Voice Transcribe, its first real-time audio perception model, capable of transcribing over 20 speakers simultaneously while switching between more than 20 validated languages, including mid-sentence code-switching. CEO Mark Zuckerberg demoed the tool on X, noting it uses adaptive delay to balance speed and accuracy. The model is now available through Meta's Mac AI app, Muse Code, and its Model API at $3 per 1,000 audio minutes.

Meta debuts Muse Voice Transcribe for real-time speech-to-text on Mac

Meta Superintelligence Labs has released Muse Voice Transcribe, a streaming speech recognition model that combines transcription, speaker diarization for over 20 voices, and endpointing in a single pass. It supports more than 70 languages, with 25 validated at launch, handles recordings over an hour long, and allows mid-sentence code-switching between languages. The model is live now in Meta AI for Mac, Muse Code, and via the Meta Model API at $3 per 1,000 audio-minutes.