Muse Voice Transcribe: Type by speaking—Meta has introduced an amazing AI tool..
New marvels are unfolding daily in the world of Artificial Intelligence (AI). Voice transcription has now become faster and more accurate than ever before. On September 1st, the tech giant Meta unveiled its new speech-to-text model, named 'Muse Voice Transcribe.' It is the first real-time audio model released by Meta Superintelligence Labs, boasting the impressive capability to instantly convert spoken conversation into text.
**Excellent Support for Five Major Indian Languages**
According to Meta, this new model supports over 70 languages. A major highlight for Indians is the inclusion of five key Indian languages: Hindi, Tamil, Telugu, Malayalam, and Kannada. In India, people often mix English with their regional languages during conversations—a practice known as 'code-switching.'
This new AI model has been designed with precisely this behavior in mind. This means that if you switch from English to Hindi mid-conversation, the tool will understand and accurately transcribe your words without any interruptions or additional processing.
**Ability to Distinguish Between More Than 20 Voices**
Another standout feature of Muse Voice Transcribe is its streaming transcription capability. This means you do not have to wait for the entire recording to finish; instead, the text appears on your screen in real time as you speak. Meta claims that the tool can distinguish between the voices of more than 20 speakers in a single audio recording.
Furthermore, it can easily transcribe recordings longer than an hour. Most notably, it requires no separate post-processing steps; a single model handles all these tasks seamlessly.
**A Perfect Blend of Speed and Accuracy**
This AI model strikes an excellent balance between speed and accuracy. It instantly transcribes simple words, while taking slightly longer to recognize complex ones. As of September 1, 2026, this tool held the top spot on the 'Artificial Analysis Streaming Speech-to-Text Leaderboard.'
How to use it and what is the cost?
Regarding availability, Muse Voice Transcribe is now accessible via Meta's model API. Currently, it is being used for dictation within Meta AI for Mac and Muse Code. The company has priced the API at $3 per 1,000 audio minutes, which works out to approximately $0.18 per hour.
At a time when the demand for speech-based AI systems for tasks like voice assistants, coding, and transcription is surging, this new move by Meta is considered highly significant. In multilingual regions like the Indian market, this tool has the potential to completely transform the way people communicate.
Disclaimer: This content has been sourced and edited from Amar Ujala. While we have made modifications for clarity and presentation, the original content belongs to its respective authors and website. We do not claim ownership of the content.

