product updateOpenAI

Pentagon Adds OpenAI's ChatGPT Mil and xAI's Grok for Government to GenAI.mil

TL;DR

The Pentagon has added OpenAI's ChatGPT Mil and xAI's Grok for Government to its GenAI.mil platform, which previously offered only Google Gemini. Anthropic's Claude remains excluded after a supply-chain risk dispute with the Trump administration.

2 min read
0

Pentagon Expands GenAI.mil With Two New Models

The US Department of Defense has added OpenAI's ChatGPT Mil and xAI's Grok for Government to its internal AI platform GenAI.mil, according to the Pentagon. Google Gemini had been the only model available on the platform since its launch in December 2025.

The Department of Defense says GenAI.mil now has more than 1.7 million users out of over three million total employees and military personnel.

Use Cases and Deployment

ChatGPT Mil is positioned mainly for administrative tasks, logistics, and planning, according to the Pentagon. Grok for Government is pitched with more military-specific framing, with the Pentagon citing procurement analysis and supply chain management as target use cases.

Both models run in a dedicated secure environment isolated from their commercial counterparts, according to the Department of Defense, which states this setup prevents sensitive government data from leaking into commercial systems. The Pentagon also says no user data is collected on the platform.

Neither OpenAI nor xAI has disclosed pricing, context window size, or specific model version details for these government-specific deployments.

Claude Remains Absent

Anthropic's Claude is not available on GenAI.mil. According to reporting, Anthropic declined to grant the Pentagon unrestricted use of its models, and the Trump administration subsequently classified the company as a supply chain risk. A court ruled that classification unlawful in late August 2026, but Claude has not yet been added to the platform.

What This Means

This is not a new model release — it's an access and deployment change for existing commercial AI systems repackaged for government use. ChatGPT Mil and Grok for Government are almost certainly wrapped versions of GPT-series and Grok models running in FedRAMP-style secure environments, not new trained checkpoints.

The more significant story here is the Claude exclusion. A court has already ruled the government's supply-chain risk designation unlawful, yet Anthropic's model still isn't on the platform a month later. That gap suggests either bureaucratic lag or continued friction over usage terms — worth watching, since Anthropic has otherwise pursued aggressive government contracting elsewhere.

The scale numbers are notable in their own right: 1.7 million active users across a platform serving over 3 million potential personnel indicates the Pentagon has moved past pilot-stage AI adoption into something closer to standard-issue tooling. Whether that adoption translates into measurable productivity gains in logistics, procurement, or planning is unverified — the Pentagon has offered use-case framing but no performance data or benchmarks tied to actual deployment outcomes.

Related Articles

product update

OpenAI Launches Agents API in Public Beta, Exposing Codex Infrastructure to Developers

OpenAI has released the Agents API in public beta, giving developers access to the same cloud infrastructure that powers Codex and ChatGPT. The API supports long-running agents, parallel tool use, and sub-agent delegation, with billing based solely on token usage.

product update

OpenAI Launches ChatGPT for Financial Services to Automate Wall Street Analyst Work

OpenAI launched ChatGPT for Financial Services, a tailored enterprise product built with design partners Morgan Stanley and Evercore that automates research, financial analysis, and pitchbook creation. The tool, powered by GPT-6 Astra, targets tasks traditionally performed by Wall Street's junior analysts and associates.

benchmark

OpenAI's GPT-6 Astra Beats Claude Fable 5.1 Nearly 3-to-1 in Autonomous Business Benchmark, Tops Drone Navigation Tests

Independent testing lab Andon Labs found OpenAI's GPT-6 Astra nearly triples Claude Fable 5.1's performance running a simulated vending machine business, averaging $15,515 versus $5,422. Astra also became the first model to beat human-AI baseline performance across all five Drone-Bench subtasks, including autonomous person-tracking via drone.

benchmark

GPT-6 Astra Beats Ai2's MolmoAct2 on New Robotics Benchmark, Researcher Calls It a 'Step Change'

A new robotics benchmark called StationeryBench shows OpenAI's GPT-6 Astra completing 7 of 100 desk-object manipulation tasks versus zero for Ai2's MolmoAct2, with a median progress score of 46 against 12. Cornell/DeepMind researcher Yoav Artzi calls the result a 'step change in spatial reasoning.'

Comments

Loading...