OpenAI Built a Hacking Prodigy, Then Locked It in a Vault

OpenAI Built a Hacking Prodigy, Then Locked It in a Vault

Imagine training the world's most gifted lockpicker, handing it a perfect score at locksmith school, and then deciding that — actually — almost nobody gets to meet it. That's roughly the energy of OpenAI's latest release: a cybersecurity model so capable the company decided you probably can't be trusted with it.

Top of the CyberGym Class

GPT-5.5-Cyber, detailed around June 24, posted an 85.6% score on the CyberGym benchmark — up from 81.8% for standard GPT-5.5 — which OpenAI says is the highest CyberGym result ever recorded by a single model. It also leapt to 39.5% on ExploitGym (from 25.95%) and 69.8% on SEC-bench Pro (from 63.1%).

Those aren't rounding-error gains. They represent a real jump in a model's ability to find and exploit software vulnerabilities — the exact skill set that keeps both red teams and the people who pay them awake at night.

A Velvet Rope Around the Cyber Brain

So OpenAI did something refreshingly grown-up: it didn't drop this thing into a public API for anyone with a credit card and a grudge. GPT-5.5-Cyber is gated behind a "Trusted Access for Cyber" program, open only to vetted organizations — think security heavyweights like Akamai, Cisco, Cloudflare, and CrowdStrike, plus government partnerships across Australia, Canada, France, Germany, Japan, and South Korea.

The logic is straightforward and a little chilling: a tool that can autonomously discover exploits is a gift to defenders and a loaded weapon for everyone else. The same model that helps Cloudflare patch a hole faster could help a ransomware crew find it first. The gatekeeping is the entire point.

It's a notable shift in tone for an industry that loves to ship first and apologize later. When even OpenAI decides a model is too spicy to open up, that's worth paying attention to.

Source: Build Fast with AI