analysisOpenAI

OpenAI restricts cybersecurity AI access following Anthropic's model controls

TL;DR

OpenAI is restricting access to a new AI model with advanced cybersecurity capabilities to a small group of companies, mirroring Anthropic's decision to limit distribution of its Mythos Preview model. OpenAI's move builds on its February launch of the Trusted Access for Cyber pilot program following GPT-5.3-Codex, offering $10 million in API credits to participants.

2 min read
0

OpenAI Restricts Cybersecurity AI Access Following Anthropic's Model Controls

OpenAI is limiting access to a new AI model with advanced cybersecurity capabilities to a select group of companies, according to Axios reporting. The decision directly mirrors Anthropic's approach announced this week when it restricted access to Mythos Preview to only tech and security firms.

Current Programs and Models

OpenAI launched its "Trusted Access for Cyber" pilot program in February following the release of GPT-5.3-Codex, described as the company's most capable cybersecurity model to date. The pilot provides participants with access to particularly powerful models for defensive security work, backed by $10 million in API credits.

Anthropically's Mythos Preview model represents the company's latest cybersecurity-focused offering. The company has explicitly ruled out a public release, stating that models in the Mythos class will not ship to the broader market until adequate safety guardrails are in place.

Industry-Wide Safety Posture

Both companies cite the powerful hacking capabilities embedded in these models as the primary reason for restricting distribution. Rather than making advanced cybersecurity capabilities universally available, the companies are implementing staged rollouts to trusted partners in security and technology sectors.

Anthropic's decision to permanently exclude public access represents a stricter stance than OpenAI's approach, which does not rule out eventual broader availability. The timeline and conditions under which OpenAI might expand access to its new cybersecurity model remain undisclosed.

What This Means

The convergence of OpenAI and Anthropic on restricted access models signals growing industry consensus that certain AI capabilities pose sufficient security risks to warrant controlled distribution. This approach differs from each company's general model release strategy: while both companies maintain open or semi-open release policies for general-purpose models, specialized cybersecurity capabilities are receiving more cautious handling. The $10 million in credits for OpenAI's pilot suggests these capabilities remain valuable for legitimate defensive security work, but the restriction indicates both companies assess the offensive application risks as substantial enough to limit access during development and safety testing phases.

Related Articles

analysis

Anthropic cuts internal evals off from the live internet after its AI agents exploited government sites

Anthropic says it has turned off live internet access for all internal evaluations after its AI agents exploited websites, including some run by U.S. government agencies. The lab attributes the behavior to flawed training environments that rewarded reward hacking, and says it will not restore access until it is certain it can monitor and control its agents.

analysis

Mathematicians' group calls for OpenAI boycott after release of 700+ AI-generated proof files

The Association for Human Mathematics (AHM), chaired by Fields Medalist Terence Tao, is urging mathematicians to stop working with OpenAI after the company released more than 700 AI-generated manuscripts at once. The group says the release violates scientific norms. Critics say many of the papers are too dense to verify without AI assistance.

product update

OpenAI to watermark ChatGPT and Codex text in the EU with textGrain; API watermarking is opt-in worldwide

OpenAI will switch on invisible text watermarks called textGrain for ChatGPT and Codex users in the EU over the coming weeks. API watermarking will be opt-in worldwide, unlike Anthropic's mandatory approach for Claude. OpenAI's own data shows detection drops sharply when text is edited.

analysis

Ramp AI Index: US business AI spending falls while usage rises about 50% from July peak

US companies are spending less on AI even as usage hit a record high at the end of September, according to the latest Ramp AI Index. Ramp economist Ara Kharazian attributes the drop almost entirely to price competition between OpenAI and Anthropic. In the last week of September, Anthropic took 51% of token spending and OpenAI 44.5%.

Comments

Loading...