cd /news/artificial-intelligence/meta-launches-ai-powered-muse-voice-… · home topics artificial-intelligence article
[ARTICLE · art-118894] src=in.mashable.com ↗ pub= topic=artificial-intelligence verified=true sentiment=↑ positive

Meta Launches AI-Powered Muse Voice Transcribe With 5 Indian Language Support

Meta launched Muse Voice Transcribe, an AI-powered real-time speech-to-text model supporting over 70 languages including Hindi, Tamil, Telugu, Malayalam, and Kannada, and topping the Artificial Analysis streaming speech-to-text leaderboard as of September 1, 2026. Developed by Meta Superintelligence Labs, the model handles speaker diarization and endpointing natively, and is available through Meta's Model API at $3 per 1,000 audio minutes. The model is already used for dictation in Meta AI for Mac and Muse Code.

read2 min views1 publishedSep 2, 2026
Meta Launches AI-Powered Muse Voice Transcribe With 5 Indian Language Support
Image: In (auto-discovered)

Meta has introduced Muse Voice Transcribe, a new AI-powered speech-to-text model designed to transcribe conversations in real time. Developed by Meta Superintelligence Labs, the model can understand and process multiple languages within the same conversation, making it useful for multilingual speakers. Muse Voice Transcribe also marks Meta’s first real-time audio perception model from the newly formed AI research division.

Supports 5 Indian Languages #

Meta says Muse Voice Transcribe can work with more than 70 languages, including Hindi, Tamil, Telugu, Malayalam and Kannada. The AI model is also built to recognise code-switching, allowing users to move between languages during the same conversation without requiring separate models.

Muse Voice Transcribe is MSL's first real-time audio perception model -- rolling out today. SOTA in streaming speech-to-text, it handles speaker diarization, and endpointing natively in a single model.

— Mark Zuckerberg (@finkd)[pic.twitter.com/LViMDSkbim][September 1, 2026]

Real-Time Multilingual Transcription #

The model generates text as people speak, rather than waiting for an entire recording to finish. It can also distinguish between speakers in recordings featuring more than 20 voices and process audio longer than an hour, with these functions handled within a single model.

Balancing Speed and Accuracy #

Muse Voice Transcribe is designed to adjust how long it listens before producing each word. It can respond quickly when speech is easy to understand while taking additional time to process words that are more difficult to recognise. Meta says the model topped the Artificial Analysis streaming speech-to-text leaderboard as of September 1, 2026, while 25 of the more than 70 supported languages had been validated at launch.

Available Through Meta’s Model API #

Meta has made Muse Voice Transcribe available through its Model API, where it costs $3 per 1,000 audio minutes, or about $0.18 per hour according to the company. The technology is already being used for dictation in Meta AI for Mac and Muse Code, highlighting its potential for transcription, coding, voice assistants and other speech-based applications.

ALSO SEE:

── more in #artificial-intelligence 4 stories · sorted by recency
── more on @meta 3 stories trending now
sponsored brought to you by zahid.host 4,200+ EU-deployed projects
reading about agents? ship yours in a single git push.

Run your AI side-project on zahid.host

EU-based hosting, git-push deploys, automatic HTTPS, no cold starts. Free tier with a custom domain — perfect for shipping the agent you just read about.

$git push zahid main
Live at https://your-agent.zahid.host
Get free account → Pricing
from €0/mo · no card required
LIVE [news/meta-launches-ai-pow…] indexed:0 read:2min 2026-09-02 ·