model releaseMicrosoft

Microsoft releases MAI-Thinking-1, its first reasoning AI model trained without third-party distillation

TL;DR

Microsoft announced MAI-Thinking-1, its first advanced reasoning AI model, at Build 2026. The company claims it's a medium-sized model matching leading models on key software engineering benchmarks, trained from scratch without distillation from third-party models.

1 min read
0

Microsoft releases MAI-Thinking-1, its first reasoning AI model trained without third-party distillation

Microsoft announced MAI-Thinking-1, its first advanced reasoning AI model, at Build 2026 on June 2. The company describes it as a "medium-sized model" that "matches leading models" on "key" software engineering benchmarks, according to Microsoft's claims.

The model represents Microsoft's most ambitious in-house development effort since introducing its initial proprietary models in 2025. Prior to that, the company relied exclusively on OpenAI's models. Microsoft and OpenAI recently renegotiated their partnership to loosen ties between the companies.

Microsoft claims MAI-Thinking-1 was "trained from the ground up on clean data, without distillation from third-party models." Specific benchmark scores, context window size, and pricing have not been disclosed.

Six additional models announced

Microsoft unveiled six other models at Build 2026:

MAI-Image 2.5 and its flash variant handle text-to-image generation and image editing.

MAI-Transcribe-1.5 is claimed to be "five times faster than competing models," though Microsoft did not specify which models it's comparing against.

MAI-Voice-2 adds 15 new languages and expanded voice options. A flash version is listed as "coming soon."

MAI-Code-1 is described as "inference-efficient" and has been integrated into GitHub Copilot and Visual Studio Code.

Microsoft did not provide technical specifications, pricing details, or release dates for general availability of any of these models.

What this means

Microsoft's announcement signals a strategic shift toward proprietary model development after years of OpenAI dependency. The emphasis on training "without distillation" suggests Microsoft is positioning MAI-Thinking-1 as a fully independent reasoning model, though without disclosed benchmarks or third-party testing, performance claims remain unverified. The simultaneous release of seven models across multiple modalities indicates Microsoft is building a comprehensive model family to compete directly with Anthropic, OpenAI, and Google—but concrete details on capabilities and availability remain limited.

Related Articles

model release

DeepSeek Releases V4-Flash-Vision-Exp, First Multimodal Model in V4 Family

DeepSeek has released DeepSeek-V4-Flash-Vision-Exp, its first experimental multimodal model in the V4 family, adding visual understanding to the V4-Flash architecture. The 305B-parameter model shows substantial gains on multimodal agent benchmarks while holding steady on text-only tasks.

model release

Tencent Open-Sources Hy4 Preview: 770B-Parameter MoE Model with 1M-Token Context

Tencent's Hy Team has open-sourced Hy4 preview, a 770-billion-parameter Mixture-of-Experts model with 49 billion activated parameters and a 1-million-token context window. The model is available under Apache 2.0 alongside an FP8-quantized variant, with Tencent claiming it beats GLM 5.3 and Kimi K3 on internal engineering evaluations.

model release

Tencent Releases Hy4 Preview: 770B-Parameter MoE Model with 1M Context for Coding Agents

Tencent has released Hy4 preview, a mixture-of-experts model with 770B total parameters and 49B active parameters, targeting coding agents and multi-step tool-use workflows. The model ships with a 1 million token context window and is priced at $0.834 per 1M input tokens and $2.501 per 1M output tokens.

model release

GLM-5.3-Flash Debuts as Zhipu AI's First Natively Multimodal Model, 320B Parameters with 18B Active

Zhipu AI has released GLM-5.3-Flash, the first natively multimodal model in its GLM-5 series, built on a 320B-parameter mixture-of-experts architecture with only 18B active parameters. The company claims it outperforms GLM-5.2 while approaching Claude Opus 4.8 on coding and agentic benchmarks at a fraction of the cost. Unsloth has published quantized GGUF versions for local inference.

Comments

Loading...

Microsoft MAI-Thinking-1: First In-House Reasoning AI Model | TPS