Meta's new real-time audio model is the foundation for AI assistants that never stop listening
The model transcribes speech, detects sentence boundaries, and distinguishes up to 20 speakers simultaneously. It processes audio in 80-millisecond chunks and costs $0.18 per hour, undercutting competitors.

Meta has released Muse Voice Transcribe, a real-time audio model designed to power AI assistants that continuously listen and respond. This model can transcribe speech, detect sentence boundaries, and identify up to 20 speakers simultaneously without requiring additional systems. Its ability to handle complex audio environments makes it a significant advancement in real-time voice processing.
The model works by breaking incoming audio into 80-millisecond chunks, allowing it to process speech in near real-time. It dynamically adjusts the delay for each word based on the difficulty of the audio, ensuring a balance between speed and accuracy. This adaptability is crucial for applications that require immediate responses, such as virtual assistants and customer service chatbots.
Meta's pricing for Muse Voice Transcribe is set at $0.18 per hour, which is significantly lower than the rates offered by competitors like OpenAI and ElevenLabs. This cost advantage positions Meta as a strong contender in the real-time audio processing market. The model also supports over 70 languages, making it a versatile tool for global applications.
The release of Muse Voice Transcribe has the potential to disrupt the AI assistant market by lowering costs and improving performance. However, it also raises concerns about data privacy and vendor lock-in, as companies may become reliant on Meta's infrastructure. The model's integration into Meta AI and the Meta Model API could further solidify Meta's influence in the AI ecosystem.
While the technology is still evolving, its current capabilities suggest a shift in how AI assistants are developed and deployed. The model's efficiency and affordability may encourage broader adoption, but its long-term impact will depend on how well it integrates with existing systems and how effectively Meta addresses potential challenges such as latency and governance.