MAI (Microsoft AI) is very far behind on coding
Summary
The article criticizes Microsoft AI's lackluster coding model performance compared to rivals like Kimi K3 and Deepseek V4, suggesting MAI is far behind despite vast resources.
Similar Articles
Microsoft tries to get back in the AI coding game with new model (1 minute read)
Microsoft plans to release a new AI coding model at its Build conference next week, aiming to regain competitiveness against rivals like Anthropic's Claude Code and OpenAI's Codex.
MAI-Thinking-1
Microsoft AI introduces MAI-Thinking-1, a 35B-active parameter reasoning model trained from scratch without distillation, achieving strong performance on software engineering and math benchmarks while emphasizing clean data and self-sufficiency.
@mustafasuleyman: Super excited to announce seven new world-class MAI models today. They represent what we consider a new era in AI desig…
Microsoft announces seven new MAI models, including a strong reasoning model (MAI-Thinking-1) and coding model (MAI-Code-1-Flash), along with Frontier Tuning for enterprise customization.
Microsoft's new MAI models
Microsoft announced two new LLMs: MAI-Thinking-1 (35B reasoning model) and MAI-Code-1-Flash (5B code model), both trained on enterprise-grade, clean data without third-party distillation, with MAI-Thinking-1 claimed to be preferred over Sonnet 4.6 in blind evaluations.
Are AI coding agents hitting a wall, or are we just measuring them wrong?
This article examines the gap between hype and reality for AI coding agents, arguing that they are effective for accelerating workflow parts but still require human oversight for architecture, debugging, and review, and questioning whether current benchmarks measure the right things.