Microsoft launches MAI-Code-1 and MAI-Thinking-1 models to reduce OpenAI dependence
Microsoft announced two proprietary AI models at its Build developer conference: MAI-Code-1 for code generation and MAI-Thinking-1 for reasoning tasks. The models are designed to run on Azure infrastructure, allowing Microsoft to reduce costs from its $13 billion OpenAI investment while competing directly with Anthropic and Google.
Microsoft launches MAI-Code-1 and MAI-Thinking-1 models to reduce OpenAI dependence
Microsoft announced two proprietary AI models at its Build developer conference in San Francisco on June 2, 2026: MAI-Code-1 for code generation and MAI-Thinking-1 for reasoning tasks.
Model specifications and availability
MAI-Code-1 is Microsoft's first coding model, designed to convert written descriptions into source code for applications and websites. The model is available through GitHub Copilot and Visual Studio Code, with Microsoft claiming it is "inference ultra-efficient." Pricing per 1M tokens has not been disclosed.
MAI-Thinking-1 is described by Microsoft as a medium-sized reasoning model "built for high efficiency and performance, but importantly, at a low-token cost," according to Kyle Daigle, Microsoft's developer marketing chief and GitHub operating chief. The model is available in private preview through Microsoft Foundry, with customers able to request testing access before broader availability.
Strategic positioning
The launch represents Microsoft's effort to establish proprietary model capabilities after investing $13 billion in OpenAI and $5 billion in Anthropic. By running models on Azure infrastructure, Microsoft can avoid paying third-party inference costs while potentially offering lower prices to developers.
The move follows Google's May 2026 announcement of Gemini 3.5 Flash, a coding and multimodal model running in Google's data centers. Microsoft is competing as both OpenAI and Anthropic pursue public offerings, with Anthropic filing confidentially for an IPO on June 1, 2026.
Additional releases
Microsoft also announced updated cloud-based models for speech recognition, synthetic voice generation, and image generation, along with small Aion models designed to run on Windows PCs. Specific model sizes, context windows, and benchmark scores were not disclosed.
What this means
Microsoft is attempting to diversify its AI strategy beyond being purely an infrastructure provider and investor. By offering proprietary models that compete with its own portfolio companies, Microsoft can control more of the value chain and potentially reduce the billions in inference costs it incurs from OpenAI and Anthropic. However, the lack of disclosed pricing and performance metrics makes it difficult to assess whether these models will meaningfully compete with established players or simply serve as cost-optimization tools for Microsoft's own products.
Related Articles
Tencent Open-Sources Hy4 Preview: 770B-Parameter MoE Model with 1M-Token Context
Tencent's Hy Team has open-sourced Hy4 preview, a 770-billion-parameter Mixture-of-Experts model with 49 billion activated parameters and a 1-million-token context window. The model is available under Apache 2.0 alongside an FP8-quantized variant, with Tencent claiming it beats GLM 5.3 and Kimi K3 on internal engineering evaluations.
Tencent Releases Hy4 Preview: 770B-Parameter MoE Model with 1M Context for Coding Agents
Tencent has released Hy4 preview, a mixture-of-experts model with 770B total parameters and 49B active parameters, targeting coding agents and multi-step tool-use workflows. The model ships with a 1 million token context window and is priced at $0.834 per 1M input tokens and $2.501 per 1M output tokens.
IBM Releases Granite 4.2 8B, a Dense Reasoning Model with 131K Context and Three Thinking Modes
IBM has released Granite 4.2 8B, a dense reasoning model built for math, code generation, and agentic workflows. The model supports 131K context, 12 languages, and three switchable reasoning modes, priced at $0.10 per 1M input tokens and $0.15 per 1M output tokens.
DeepSeek Releases V4-Flash-Vision-Exp, First Multimodal Model in V4 Family
DeepSeek has released DeepSeek-V4-Flash-Vision-Exp, its first experimental multimodal model in the V4 family, adding visual understanding to the V4-Flash architecture. The 305B-parameter model shows substantial gains on multimodal agent benchmarks while holding steady on text-only tasks.
Comments
Loading...