Register and share your invite link to earn from video plays and referrals.

Mustafa Suleyman
@mustafasuleyman
CEO, @MicrosoftAI | Author: The Coming Wave | Past: Co-founder, @InflectionAI & @GoogleDeepMind
494 Following    936.5K Followers
Our first reasoning model, MAI-Thinking-1, is built from scratch. Now available in Microsoft Foundry. Kudos to the team! More below.
We believe this is a gamechanger, and our models and agentic harnesses are only going to get better. Learn more here and sign up!
Big news! Our new MAI-Cyber-1-Flash model combined with MDASH, our multi agent security harness, delivers 96% on the CyberGym benchmark, 12pts above Mythos, at HALF the cost. Proud of the team. More details in THREAD:
Show more
0
52
968
108
Forward to community
As @satyanadella says, we're making great progress on shipping MAI models that are higher quality, faster and cheaper across MSFT. In PowerPoint, our image model cut costs 84% compared with GPT-Image-2. In OneDrive, it lifted save rates 26% and cut latency by ~25%. It's also the default model in Bing delivering great quality and performance. In Dragon Copilot, our transcription model now covers 58 languages and halves the error rate on multilingual clinical transcription. And of course the flights we have in motion on GitHub and Excel will no doubt be huge too! More details in the blog here:
Show more
Introducing Ode Poetry. Ode is a wonderful poetry pharmacy that reads you a poem for the moment you’re in. Just tell Ode what you're feeling, and it uses Microsoft AI audio models to connect you with the same work that poetry expert William Sieghart would recommend. The best technology doesn't replace human creativity, it helps more people experience it. Super proud of the team for making this truly humanist tool. More in the blog:
Show more
0
170
439
40
Forward to community
Super excited to announce seven new world-class MAI models today. They represent what we consider a new era in AI designed to keep you in control and on the frontier. First is our text foundation model, MAI-Thinking-1, exceptionally strong on reasoning and SWE tasks. - It’s a 35B active parameter MoE with a 256K context window. Independent human raters on Surge prefer it for overall quality in blind side-by-sides versus Sonnet 4.6, and it’s achieved 97% on AIME 2025, the key measure of its general-purpose reasoning abilities. - It's at 53% on SWE Bench Pro, placing it right alongside Opus 4.6 on one of the toughest coding benchmarks. - And since we co-designed our models with our own silicon, MAI-Thinking-1 is optimized on our MAIA 200 chip. Benchmarking head-to-head against the GB200, we see 30% better performance per dollar as well as a 1.4x performance-per-watt gain when running our MAI models on the MAIA 200 end-to-end. Next is MAI-Image-2.5 and its Flash variant. Two super strong models now at #2# on the leaderboards, surpassing the score of Nano Banana 2 on image editing. Last for now is MAI-Code-1-Flash, our new inference efficient coding model, especially tuned for VS Code and GitHub Copilot CLI. - Code-1-Flash achieves 51% on SWE Bench Pro, despite having just 5B parameters, putting it closer to Haiku in size but cheaper in cost. All of this is the foundation for Microsoft Frontier Tuning. It lets you customize our models to create custom, company-specific agents that only you control. You can make our model, your model. Your data. Your agents. Your moat. Early adopters are already seeing a difference. When we tuned our models for McKinsey’s tasks, MAI delivered the highest win rate, outperforming GPT-5.5 on quality, while being 10x lower on cost. Also really excited to be collaborating with the amazing team at Mayo Clinic to jointly train a new frontier AI model for healthcare. Our announcements today mark another milestone on the road to humanist superintelligence. You can learn more and about our other new models in our latest blog:
Show more
0
192
3.8K
541
Forward to community