Meta Releases Real-Time Transcription Model Muse Voice Transcribe
First seen · 9/6/2026, 05:45 PMLatest activity · 9/6/2026, 05:45 PM
Meta has introduced Muse Voice Transcribe, a streaming audio model capable of parsing spoken input in 80-millisecond chunks while simultaneously handling speaker diarization and sentence boundary detection. Benchmark data from Artificial Analysis positions the system as both the most accurate and lowest-cost streaming option on the market. Designed with hardware like smart glasses in mind, the release provides the essential infrastructure for continuous, ambient conversational agents.
Event heat · last 24 hours
There are 8 persisted snapshots in the last 24 hours. Peak heat was 10 at 9/12, 14:00; latest heat is 10.