@MicrosoftAI

Building a new class of safer, more capable AI systems we call Humanist Superintelligence: AI that is always aligned, controllable, and in service of humanity.

Joined July 2017
MAI-Code-1.1-Flash is going local with on-device model calls inside Github Copilot. High-quality agentic coding on your device, with no inference charge for local model calls.
The next era of the PC is taking shape on Windows. ✨ Windows is becoming the home for hybrid intelligence. 🔐 MXC GA 🧠 Frontier coding models running locally ✨ GitHub Copilot HydraFusion 💻 llama.cpp + Windows ML ⚡ Surface Laptop Ultra + Dev Box preorders 🖥️ New RTX Spark PCs
62
224
55
1,944
112,691
Microsoft AI retweeted
Listen to this support call using @MicrosoftAI's speech models and LiveKit Agents. • MAI-Transcribe-2-Streaming hears you • Gemma 4 on LiveKit Inference reasons and calls tools • MAI-Voice-2.1-Flash answers What do you think? STT: livek.it/E4qoZvU TTS: livek.it/RsCc37l
5
5
1
78
5,287
Build expressive voice agents with MAI models now in LiveKit. Learn more here: msft.it/6014alofr
MAI-Transcribe-2-Streaming, MAI-Voice-2.1, and MAI-Voice-2.1-Flash from @MicrosoftAI are now live on LiveKit. Transcribe-2-Streaming debuts at #1 on the Artificial Analysis accuracy leaderboard → Pair it with MAI-Voice-2.1-Flash to build efficient and expressive voice agents. Try them out: docs.livekit.io/agents/model…
4
4
89
7,537
More ways to discover, access, and build with MAI models. MAI models are now available through Vercel, giving developers another way to bring Microsoft AI models into the products and experiences they’re building. This includes our existing MAI models plus today’s newest releases: MAI-Transcribe-2-Streaming, MAI-Voice-2.1, and MAI-Voice-2.1-Flash.
We partnered with @MicrosoftAI to bring their secure, transparent, and cost-efficient models to AI Gateway. Available today: • 𝙼𝙰𝙸-𝚅𝚘𝚒𝚌𝚎-𝟸.𝟷: long-form & fast-reply speech • 𝙼𝙰𝙸-𝚃𝚛𝚊𝚗𝚜𝚌𝚛𝚒𝚋𝚎-𝟸-𝚂𝚝𝚛𝚎𝚊𝚖𝚒𝚗𝚐: transcription vercel.com/changelog/microso…
10
3
133
11,738
Introducing 3 new models: MAI-Transcribe-2-Streaming, MAI-Voice-2.1 and MAI-Voice-2.1-Flash. Accurate streaming transcription. Natural speech and less waiting between turns. Build voice agents that keep the conversation moving!
97
160
56
1,607
142,272
MAI is hill-climbing 61% faster than any other lab in image generation. We've gained +283 Elo on text-to-image in 10 months to optimise the pareto frontier.
28
24
7
311
26,089
No lab has reached the image frontier faster. Microsoft AI: 864 → 1147 Elo in nine months. That's 377 points a year — 1.6× the next-fastest climb on the board.
4
3
35
2,484
Microsoft AI retweeted
More exciting news from @MicrosoftAI: MAI-Image-2.6 is #2 in Text-to-Image Arena with 1,331 pts, and joins the Pareto frontier! At $38.90 per 1K output images (~$0.039/image), it outperforms Grok Imagine Image 2.0 while also costing less. MAI-Image-2.6 ranks #2 in: - Product, Branding & Commercial Design - 3D Imaging & Modeling - Cartoon, Anime & Fantasy - Art - Text Rendering It also ranks #3 in Portraits and Photorealistic & Cinematic Imagery. Try it in Direct Mode, or vote in Battle Mode to help shape the live leaderboard.
MAI-Image-2.6 by @MicrosoftAI has landed in the Text-to-Image Arena at #2 (1336 pts)! This release is just 45 pts behind GPT Image 2 (Medium) at #1, and ahead of #3 Grok Imagine Image 2.0 (Low) by 20 pts. MAI-Image-2.6 is a significant improvement from MAI-Image-2.5 overall at #10 with (1256 pts), and across all categories. MAI will also be making it available on Playground with early API access on Foundry next week. See post below for more details. Congrats to the @MicrosoftAI team on this release!
7
12
2
208
28,349
Microsoft's MAI-Image-2.6-Flash takes #3 on the Artificial Analysis Image Editing Leaderboard, a significant jump over the previous generation’s Flash variant and joining MAI-Image-2.6 on the Pareto frontier for quality vs price MAI-Image-2.6-Flash is an optimized version of MAI-Image-2.6, Microsoft AI's flagship image model, released today on Microsoft Foundry alongside MAI-Image-2.6. Like the rest of the family, it handles both text to image generation and image editing. In the Artificial Analysis Image Arena, MAI-Image-2.6-Flash lands at #3 in Image Editing, behind only Microsoft's own MAI-Image-2.6 and OpenAI's GPT Image 2 (high) and narrowly ahead of Google's Nano Banana 2. In Text to Image it takes #8, within 5 Elo points of MAI-Image-2.5 and narrowly ahead of Google's Nano Banana Pro. MAI-Image-2.6-Flash is a large step up on MAI-Image-2.5-Flash at the same price: it sits 69 Elo points higher in Text to Image (#16 to #8) and 34 Elo points higher in Image Editing (#12 to #3). MAI-Image-2.6 and MAI-Image-2.6-Flash are available today on Microsoft Foundry. Congratulations to @MicrosoftAI on the release! See below for our analysis and example outputs of MAI-Image-2.6-Flash in the Artificial Analysis Image Arena 🧵
17
20
5
322
40,979
MAI-Image-2.6-Flash has the best price/quality in the world.
39
36
12
523
28,141
MAI-Image-2.6 and new MAI-Image-2.6-Flash are redefining the Artificial Analysis quality/price Pareto frontier for image editing. Both now in Microsoft Foundry: msft.it/6018aVgIQ
8
1
1
35
4,579
Thanks for the support @AndrewYNg! Completely agree, faster token generation will become increasingly important as a greater proportion of output tokens are consumed by models, such as in multi-step agentic workflows, rather than being read by people.
Shoutout to the team that built artificialanalysis.ai/ . Really neat site that benchmarks the speed of different LLM API providers to help developers pick which models to use. This nicely complements the LMSYS Chatbot Arena, Hugging Face open LLM leaderboards and Stanford's HELM that focus more on the quality of the outputs. I hope benchmarks like this encourage more providers to work on fast token generation, which is critical for agentic workflows!
63
53
7
850
844,268
MAI-Transcribe-2: the highest quality, cheapest transcription at the fastest speed! 10x faster that GPT-Transcribe. Now available on Microsoft Foundry.
28
58
29
605
45,702