In a fresh development, this doesn't affect our editorial independence. When you purchase through links in our articles, we may earn a small commission.
Industry observers note that less than a week after touting the scientific achievements of Astra, its next “major” model, OpenAI says it’s “pausing internal activities” related to the model due to its powerful cybersecurity abilities.
In a fresh development, “Our most recent internal evaluations of Astra, one of our upcoming models, over the past few days indicate significant advancements in agentic coding and cybersecurity,” OpenAI stated in a Friday press drop.
Industry observers note that “These results, in addition to expert assessments, have led us to conclude last night that we cannot rule out critical cyber capabilities under our Preparedness Framework.”.
As part of the ongoing story, openAI’s Preparedness Framework outlines scenarios in which development of a fresh model should “halt” if it reaches certain capability thresholds in various categories, including “Biological,” “Cybersecurity,” and “AI Self-improvement.”.
Industry observers note that for cybersecurity, the “critical” threshold means a model can pinpoint “zero-day exploits of all severity levels” in “hardened real-global stage systems” without any human help.
Industry observers note that a model could also hit the “critical” level if it can carry out “end-to-end novel strategies for cyberattacks against hardened targets” with little more than a “high-level desired goal” in mind, according to the OpenAI safety framework.
The report highlights that openAI initially dropped GPT-5.6 Sol to just a “select group of trusted partners” before making the model public a couple of weeks later. OpenAI’s previous high-end model, GPT-5.6 Sol, only reached the “high” threshold during internal evaluations, the publisher said.
In a fresh development, given its concerns over Astra’s potential cybersecurity risks, OpenAI says is it “implementing stricter security controls” for the model, such as setting up “isolated testing environments” and “restricted network and tool access,” among other measures.
In a fresh development, in the meantime, OpenAI is “pausing internal activities involving Astra that do not yet meet these strengthened security control requirements,” the publisher said.
As part of the ongoing story, openAI said it sounded its warning about Astra because “it’s important to be transparent to the public” about what Astra is potentially capable of.
In a fresh development, barely a week ago, OpenAI touted Astra’s abilities in mathematical research, including its solutions to 10 open math and computer science problems.
Industry observers note that openAI’s revelations about Astra come amid a flurry of reports of advanced AI models going rogue, hacking real firms and organizations during training exercises and even forging phony credentials to hack external systems.
The report highlights that now with Astra said to be demonstrating dangerously strong cybersecurity capabilities, it seems we may have reached a crossroads in AI safety, where each fresh “frontier” model on the AI test bench is judged — at least initially — too powerful to be dropped.
As part of the ongoing story, his coverage of artificial intelligence interrogates the most recent LLMs, and how they can be used at work and at home to be best prepared for the AI revolution. “AI is going to change our lives sooner than we think,” Ben writes. “Our best way to adapt is by using it every day.” Ben has been a PCWorld author since 2014, and has covered everything from laptops to security cameras before launching PCWorld’s AI beat. Ben's articles have also appeared in PC Magazine, TIME, Wired, CNET, Men's Fitness, Mobile Magazine, and more. Ben holds a master's degree in English literature. Ben has been writing about consumer technology for more than 20 years, and now focuses his reporting on AI as it relates to the basic human experience.