model releaseMicrosoft

Microsoft releases MAI-Thinking-1, its first reasoning model with 35B parameters

TL;DR

Microsoft released seven AI models at Build 2026, headlined by MAI-Thinking-1, its first reasoning model with 35 billion parameters. The company claims the model matches Anthropic's Claude Opus 4.6 on SWE Bench Pro coding benchmarks and beats Sonnet 4.61 in blind tests.

2 min read
0

Microsoft releases MAI-Thinking-1, its first reasoning model with 35B parameters

Microsoft launched seven AI models at its Build 2026 developer conference, including MAI-Thinking-1, the company's first reasoning model.

MAI-Thinking-1 specifications

The 35-billion-parameter model was trained on "enterprise-grade, clean and commercially licensed data," according to Microsoft. The company claims it beat Anthropic's Claude Sonnet 4.61 when evaluated by independent reviewers in blind tests, and matches Claude Opus 4.6's SWE Bench Pro benchmark score for coding. Specific benchmark numbers were not disclosed.

MAI-Thinking-1 is designed for multi-step agentic tasks and is available in private preview through Microsoft Foundry. Context window size and pricing have not been announced.

Six additional models

MAI-Code-1: A coding model Microsoft describes as "ultra-efficient" and "tuned for GitHub." Available today in Copilot and VS Code. No parameter count or benchmarks disclosed.

MAI-Image-2.5 and MAI-Image-2.5 Flash: Microsoft's first text-to-image and image-to-image models. According to Microsoft, MAI-Image-2.5 outperformed Nano Banana Pro on ELO ratings and ranked third on the LM Arena Leaderboard at launch, behind Nano Banana. Available now in PowerPoint, Foundry, and rolling out in OneDrive.

MAI-Transcribe-1.5: A speech transcription model supporting 43 languages, with streaming capabilities coming soon. This represents an update to the MAI-Transcribe line released just two months prior.

MAI-Voice-2 and MAI-Voice-2 Flash: Text-to-speech models supporting 15 more languages than the previous MAI-Voice-1, which was released in preview two months ago.

Microsoft AI CEO Mustafa Suleyman stated that all models include watermarking "from scratch" and offer cost efficiency improvements up to 10x compared to competitor models, though specific pricing was not provided.

Availability and healthcare partnership

All MAI models will be available through Fireworks AI (now generally available on Foundry), Baseten, and OpenRouter.

Microsoft also announced a partnership with Mayo Clinic to develop a frontier model for healthcare, joining existing medical AI efforts from OpenAI and Google. Microsoft already offers Copilot Health.

What this means

Microsoft's entry into reasoning models puts it in direct competition with OpenAI's o1 series, Anthropic's Claude family, and DeepSeek's R1. The emphasis on commercially licensed training data addresses enterprise customers' copyright concerns, though the lack of disclosed pricing and limited benchmark data makes performance comparisons difficult. The rapid release cycle—with voice and transcription model updates arriving just two months after initial previews—signals Microsoft's aggressive push to compete across the full AI model stack.

Related Articles

model release

OpenAI Halts Parts of Astra Model Development After It Hit 'Critical' Cybersecurity Threshold

OpenAI disclosed that its in-development Astra model showed cyberattack capabilities strong enough that it cannot rule out a 'Critical' risk classification. The company has paused related internal activity and added security controls under its Preparedness Framework.

model release

Mistral's 3B-Parameter Shieldstral Matches 20B Safety Model on Text Benchmarks

Mistral's new Shieldstral, a 3-billion-parameter open-weight safety classifier, posts an 84.9% F1 score on text benchmarks—tying OpenAI's GPT-OSS-Safeguard-20B, a model roughly seven times larger. The model lets operators define safety rules at runtime using plain-language yes/no questions instead of fixed taxonomies.

model release

Mistral AI Releases Shieldstral-1.0-3B, a 3B-Parameter Policy-Adaptive Safety Classifier

Mistral AI has released Shieldstral-1.0-3B, a compact open-weight safety classifier that evaluates text and images against natural-language policies specified at inference time. The 3B model runs on a single GPU and reports F1 scores competitive with or exceeding larger moderation models like LlamaGuard-4-12B and GPT-OSS-Safeguard-20B on multiple benchmarks.

model release

Black Forest Labs Launches FLUX 3 Video, Claims It Beats Seedance 2.0 on Elo Rankings

Black Forest Labs has made FLUX 3 Video generally available via its API, offering up to 20-second HD/Full HD clips with native audio and lip-sync in 14+ languages. The company claims its internal Elo benchmarks put the model ahead of Seedance 2.0, Gemini Omni Flash, and Minimax H3.

Comments

Loading...