Multilingual transcription
Supported speech recognition across 25 languages.
Historical tool archive
MAI-Transcribe-1 was Microsoft's batch-oriented speech-to-text model for 25 languages; Microsoft deprecated it on August 20, 2026 in favor of MAI-Transcribe-1.5.
No longer active

Archive
MAI-Transcribe-1 launched in April 2026 as a multilingual speech-recognition model in Microsoft Foundry and Azure Speech. It supported 25 languages, accepted WAV, MP3, and FLAC audio, and was priced at $0.36 per hour of audio.
Microsoft deprecated MAI-Transcribe-1 on August 20, 2026. New implementations should use MAI-Transcribe-1.5, which expands support to 43 languages and adds streaming, phrase biasing, and transcript-style controls.
Past capabilities
Supported speech recognition across 25 languages.
Was designed for fast transcription of uploaded audio.
Accepted WAV, MP3, and FLAC inputs through Azure Speech.
Microsoft designed the model for accents, background noise, and imperfect recordings.
Availability
Historical pricing was $0.36 per hour of audio. MAI-Transcribe-1 is deprecated, so use the current successor's official pricing for new workloads.
Status checked . Review the historical source ↗
Current options
Content Creator
Consider Google's current transcription model for a maintained cloud alternative.
Explore Gemini 3.5 Transcribe →Miscellaneous
Consider Voxtral Transcribe 2 for multilingual transcription and a realtime option.
Explore Voxtral Transcribe 2 →Consumer
Consider Cohere Transcribe when an open-source speech-recognition model better fits the deployment.
Explore Cohere Transcribe →Questions
No. Microsoft documents MAI-Transcribe-1 as deprecated on August 20, 2026.
MAI-Transcribe-1.5 is the direct successor. It expands language support and adds features such as streaming, phrase biasing, and transcript-style selection.
Microsoft published a price of $0.36 per hour of audio for MAI-Transcribe-1.

Critique - Microsoft's multi-model deep research tool that pits AI models against each other

Google Edge Eloquent - Google AI Edge Eloquent — Free voice dictation app that turns messy speech into polished text, runs fully offline, no subscription, no usage caps

TADA - Hume AI's open-source TTS model that syncs text and audio one-to-one for zero-hallucination speech

Harrier - Microsoft Bing’s SOTA, open-source embedding model for search and RAG grounding

Get access to all our AI courses, hundreds of real-world AI use cases, live expert-led workshops, an exclusive network of AI early adopters, and more.
Get unlimited access to all of our current & upcoming industry-specific AI courses for the duration of your subscription.
To keep up with the rapid pace of AI, our team publishes AI implementation guides daily. Our library contains 300+ practical use cases to automate real-world work.
Join weekly, live, interactive sessions with industry leaders who are at the forefront of AI for hands-on implementation guidance and exclusive insights.
Network with an exclusive community of AI-first professionals who are working smarter with AI. Learn how early adopters are using AI in their work and businesses.