← Back to explorer

MAI-Voice-2.1-Flash

Type
model
Venue
Microsoft AI announcement, Oct 1, 2026
Year
2026
Source
blog
Access
paid
Language
English
Added
2026-10-01
Verified
2026-10-01

Summary

Faster variant of MAI-Voice-2.1 with the same languages and voices: 55% faster inference, ~60% cheaper than comparable models.

Keywords

text-to-speech · TTS · voice agents · multilingual · fast inference

Topics

speech, text-to-speech, voice agents

Research notes

  • Discovery: @MicrosoftAI X post 2026-10-01 (https://x.com/MicrosoftAI/status/2105693024013467905)
  • Blog: https://microsoft.ai/news/our-first-streaming-transcription-model/
  • Demo: Chatter live voice-agent demo at http://playground.microsoft.ai
  • Access: OpenRouter (https://aka.ms/mai-openrouter), Microsoft Foundry, Vercel (https://aka.ms/mai-vercel), LiveKit (coming soon), Azure Voice Live
  • No Hugging Face/arXiv/GitHub release listed