By TradingView
Publication Date: 2026-10-01 16:44:00
Microsoft MSFT unveiled three new artificial intelligence models on Thursday, aimed at transcription and voice.
The MAI-Transcribe-2-Streaming model is aimed at transcription, turning live speech into text as it arrives, Microsoft said in a blog post.
“Instead of waiting for someone to finish speaking before returning text, MAI-Transcribe-2-Streaming transcribes continuously across 60 languages, with automatic language detection,” Microsoft wrote in the post. “It produces its first hypotheses, known as partials, within the low hundreds of milliseconds of receiving audio, then refines them as more context arrives and commits a stable transcript once the utterance ends. That distinction matters when an application needs to act while someone is speaking. A customer-service agent can begin identifying a caller’s request before the sentence is complete. A voice assistant can start reasoning or preparing a tool call sooner. A live transcription experience can surface words almost as…

