September 28, 2026Updated daily by the AI editorial team
← 🤖 AI

2026-09-02

OpenAI’s ‘Astra’ Hits a Critical Cyber Threshold—and Triggers a Safety Rethink

On September 1, OpenAI said its upcoming frontier model “Astra” is so capable in cybersecurity that it must be treated as a special case. Internal evaluations found that Astra can automatically identify more software vulnerabilities than any OpenAI model available to the public today, leading the company to classify it as the first system to reach a “critical” cyber capability threshold.

Because of that designation, OpenAI is delaying broad release. Astra will initially roll out only to selected partners and vetted researchers, while engineers add additional guardrails to prevent abuse for offensive hacking. The company previously paused some training runs and cyber evaluations to reassess how to test and deploy the model more safely.

Those restrictions come with an obvious tradeoff. Limiting Astra’s capabilities will also constrain legitimate uses, such as helping security teams find and fix bugs before attackers do. The episode illustrates a growing tension in AI: as models become powerful tools for both defense and offense in cyberspace, companies and regulators are being forced to define where the line for acceptable risk should be drawn.

Source: OpenAI to limit access to Astra's most powerful cyber tools