product updateOpenAI

OpenAI releases GPT-5.5-Cyber with 85.6% CyberGym score, surpassing restricted Anthropic model

TL;DR

OpenAI released an updated GPT-5.5-Cyber model that scores 85.6% on CyberGym, surpassing Anthropic's Mythos 5 (83.8%) — the same model that triggered Trump administration export controls. The release proceeds without the political pushback that forced Anthropic to restrict foreign national access.

2 min read
1

OpenAI released an updated version of GPT-5.5-Cyber on Monday that achieves an 85.6% score on CyberGym, an internal benchmark measuring AI agents' ability to reproduce known software vulnerabilities. The model's capabilities exceed those of Anthropic's Mythos 5, which scored 83.8% on the same evaluation according to Anthropic's system card.

The release raises questions about the Trump administration's selective enforcement of AI security concerns. While Anthropic faces export controls barring foreign nationals from accessing Fable 5 and Mythos 5, OpenAI's more capable cybersecurity model deployed without apparent restrictions or political intervention.

Diverging regulatory treatment

OpenAI announced the GPT-5.5-Cyber update alongside expanded partnerships with organizations in Australia, Canada, France, Germany, Japan, Poland, South Korea, and the EU. The company did not respond to requests for comment about coordination with federal authorities.

In contrast, negotiating access to Mythos dominated discussions at last week's G7 Summit. Anthropic remains subject to export directives that restrict international use of its models, despite achieving lower benchmark scores than OpenAI's newly released system.

The White House did not respond to requests for comment on the apparent inconsistency in treatment between the two companies.

CyberGym benchmark context

CyberGym measures whether AI agents can successfully reproduce known software vulnerabilities — a capability that raises both defensive and offensive security concerns. The 1.8 percentage point difference between GPT-5.5-Cyber (85.6%) and Mythos 5 (83.8%) represents a meaningful performance gap on this evaluation.

OpenAI positioned the release within a broader cybersecurity initiative, announcing partnerships with security companies and researchers. The company did not disclose specific technical changes from previous GPT-5.5-Cyber versions.

Political dimensions

Reports suggest personality conflicts between Anthropic leadership and the Trump administration contributed to the export controls, beyond purely technical security assessments. The ability of OpenAI to deploy a more capable cybersecurity model without similar restrictions suggests non-technical factors influenced the regulatory divergence.

The situation has created operational challenges for cybersecurity defenders who rely on advanced AI models, with some organizations unable to access Anthropic's restricted systems despite their defensive use cases.

What this means

The inconsistent application of AI security controls between comparable models from different companies signals either incomplete threat assessments or politically-influenced regulation. Organizations building cybersecurity defenses now face uncertainty about which capabilities will remain accessible and under what conditions. The CyberGym benchmark scores provide quantifiable evidence that regulatory restrictions did not correlate with demonstrated technical capabilities.

Related Articles

changelog

OpenAI Cuts GPT-6 Sol and Luna Prices in Half, but Independent Benchmarks Show Flat Performance

OpenAI's GPT-6 Sol and Luna cut input/output token prices in half versus GPT-5.6, with Sol now at $2/$10 per million tokens and Luna at $0.10/$0.50. Independent testing from Artificial Analysis shows intelligence scores barely moved, with regressions on some knowledge-work benchmarks.

model release

Anthropic and OpenAI Cut Prices With Claude Opus 5.5, GPT-6 Sol and GPT-6 Luna

Anthropic released Claude Opus 5.5, claiming roughly 40% lower running costs than Opus 5, while OpenAI introduced GPT-6 Sol and GPT-6 Luna with API prices cut 50% from GPT-5.6 promotional rates. The releases mark the first launches from either lab since Anthropic CEO Dario Amodei called for an industry slowdown on advanced AI development.

model release

OpenAI Releases GPT-6 Sol and Luna at Half the Price of Predecessors

OpenAI has released GPT-6 Sol and Luna, updated versions of its mid-tier and lightweight models, priced at half the cost of their GPT-5.6 predecessors. The company claims GPT-6 Sol makes roughly half as many factual errors as its predecessor, reaching what it calls 'Astra-level reliability' at lower cost.

changelog

OpenAI Python SDK v3.19.0 Adds GCP Storage Support and References Unreleased 'GPT-Rosalind' Model

OpenAI released v3.19.0 of its Python SDK on September 22, 2026, adding GCP external storage support and a code reference to an unannounced research model called GPT-Rosalind. The release also ships five bug fixes covering WebSocket handling, retry logic, and async compatibility.

Comments

Loading...