Meta Launches Multilingual Voice Tool to Enable Transcription in Five Indian Languages

US-based technology giant Meta has introduced its first real‑time audio‑perception model -- Muse Voice Transcribe -- which offers native support for five major Indian languages and advanced streaming features.
Muse Voice Transcribe -- developed by Meta Superintelligence Labs -- delivers streaming transcription, speaker separation across over 20 voices in hour‑plus recordings, native code‑switching and diarisation from a single model with no post‑processing step, the company said in a statement.

Muse Voice Transcribe is trained across over 70 languages spoken in multiple countries with 25 validated at launch.
"It ranks first on the Artificial Analysis streaming speech-to-text leaderboard as of September 1, 2026," the statement added.
Available via the Meta Model API and already running dictation in Meta AI for Mac and Muse Code, Muse Voice Transcribe delivers real‑time automatic speech recognition, diarisation with over 20 speakers and endpointing.
It is multilingual with seamless code-switching and improves accuracy with language, keyword, and context biasing.
“The longer the model waits to predict, the more accurate the transcript, but the higher the latency. Muse Voice Transcribe has “adaptive delay,” dynamically changing delay for each word based on difficulty,” the statement noted.
This is enabled with reinforcement learning (RL), where word error rate (WER) reward and a delay reward are combined multiplicatively.
“Muse Voice Transcribe is an autoregressive multimodal model from the Muse Spark family,” Meta said.
It explained that the audio is processed in 80 ms chunks (12.5 Hz), each of which is transformed into a single soft token. At each audio chunk, the model decides to either continue listening to the next audio chunk or emit a text token.
With adaptive delay, Muse Voice Transcribe achieves the Pareto front on speed-accuracy trade-off measured by time to final transcription, the company said.

Shilpa Shetty urges fans to ‘pause, notice and appreciate’ nature on World Environmental Health Day

India’s bond market performs well, equities may regain attractiveness: RBI’s Poonam Gupta

International flight passengers in South Korea rise 9.8% to 68.44 million in Jan-Aug

India urges fair pathogen access and benefit sharing under WHO Pandemic Agreement

Chiranjeevi’s ‘Kaaka’ Begins Dubbing; Film Set for January 13, 2027 Release

Odisha textbooks had 1,741 errors after ₹175 crore printing spend

Meghalaya to introduce AI-powered learning tools in 31 government, aided schools

Karnataka minister seeks relief for teachers from SIR duties, plans Bill on non-teaching work

E-Jaadui Pitara: NCERT’s digital learning app blends AI with stories, games and activities

9-year-old memorises 4,000 Panini sutras, builds offline Gita app

Shilpa Shetty urges fans to ‘pause, notice and appreciate’ nature on World Environmental Health Day

India’s bond market performs well, equities may regain attractiveness: RBI’s Poonam Gupta

International flight passengers in South Korea rise 9.8% to 68.44 million in Jan-Aug

India urges fair pathogen access and benefit sharing under WHO Pandemic Agreement

Chiranjeevi’s ‘Kaaka’ Begins Dubbing; Film Set for January 13, 2027 Release

Odisha textbooks had 1,741 errors after ₹175 crore printing spend

Meghalaya to introduce AI-powered learning tools in 31 government, aided schools

Karnataka minister seeks relief for teachers from SIR duties, plans Bill on non-teaching work

E-Jaadui Pitara: NCERT’s digital learning app blends AI with stories, games and activities

9-year-old memorises 4,000 Panini sutras, builds offline Gita app
Copyright© educationpost.in 2024 All Rights Reserved.
Designed and Developed by @Pyndertech