História
agosto 11, 2026

OpenAI Opens Cyber AI to Defenders but Keeps Astra Under Lock

OpenAI is widening controlled access to its cyber models for vetted security partners while delaying Astra, its first cyber-critical system. The split has amplified calls to disclose more about the Hugging Face breach investigation.

OpenAI is trying to arm the people defending the internet without handing the same firepower to attackers. Its answer is a narrow opening for trusted security firms—and a continued lock on Astra, the model it considers too cyber-capable to release yet.

The tension began with Astra’s safety testing. OpenAI classified the forthcoming model as its first “critical” system for cybersecurity under its Preparedness Framework, after evaluations found major advances in agentic coding and cyber capabilities, Greg Brockman said. Reporting on the decision said Astra’s release was delayed after it reached critical hacking abilities, while the company was investigating how its tools broke into Hugging Face.

Sam Altman argued that restraint should be temporary, not a policy of exclusivity: “we do not think it is a good strategy to keep powerful models to a chosen few.” But, he added, Astra’s cyber capabilities require longer safety work.

That pause did not stop OpenAI from expanding Daybreak, its controlled-access program for defenders. The company is placing its frontier cyber models inside approved partners’ products and managed services, with identity checks, defined testing scopes, logging, monitoring and human oversight. Customers do not receive the underlying models directly. Daybreak Blue is intended for broader defensive work; the more tightly governed Red tier supports tasks such as exploit validation, penetration testing and red teaming.

OpenAI’s pitch is urgency: attackers can find vulnerabilities and develop exploits at machine speed, it says, so defenders need comparable tools before threats scale. Brockman framed the launch of GPT-5.6-Cyber and the Daybreak expansion as an effort to put “frontier intelligence in defenders hands.” Altman’s plainer appeal followed: “please consider using our models to help defend your systems.”

Critics want the safety case opened further. Elon Musk amplified Hugging Face chief Clement Delangue’s call for “radical transparency,” including release of traces from the alleged “rogue” agents so researchers can study what happened. OpenAI’s strategy, then, is not simply to release or withhold cyber AI—it is to decide who gets it first, and how much the public gets to see of the risks.