OpenAI Delays Next Major AI Model 'Astra' Over Critical Hacking Concerns - MacRumors
Skip to Content

OpenAI Delays Next Major AI Model 'Astra' Over Critical Hacking Concerns

OpenAI today said it is "pausing" activities involving its upcoming AI model Astra, because its cyber capabilities are potentially too dangerous. OpenAI says its newest internal evaluations show "significant advancements in agentic coding and cybersecurity," and it cannot rule out "critical cyber capabilities." Prior OpenAI models, including GPT–5.6 Sol, were labeled as "High."

openai logo word mark
Astra triggers stricter guidelines in OpenAI's "Preparedness Framework." The guidelines call for caution when developing frontier AI capabilities that create risks of severe harm, and the cybersecurity portion of the framework says OpenAI will implement extra safeguards for models that "create new risks of scaled cyberattacks and vulnerability exploitation."

The "Critical" threshold Astra may have hit is defined by an ability to identify and develop functional zero-day exploits of all severity levels in many hardened real-world critical systems without human intervention, or devise and execute end-to-end novel strategies for cyberattacks.

OpenAI says it is increasing its safeguards and security controls before deploying Astra, including limiting work on the model until new safeguards are in place. The company plans to use isolated testing environments with restricted network and tool access, along with adding sandboxed execution and more monitoring capabilities. OpenAI says it will work with relevant government agencies and AI safety organizations to test Astra.

"We're committed to working alongside governments, safety institutes, and civil society to ensure that the frontier capabilities of models like Astra, and those that follow, are deployed responsibly and broadly for the benefit of all humanity," writes OpenAI.

Astra wasn't formally announced, but OpenAI shared details on its next major model in a recent post outlining its mathematical advancements. Astra solved 10 open problems in math and theoretical computer science for around $2,000 (in Sol API rates).

Advancements in AI are changing cybersecurity for major tech companies like Apple by unearthing an unprecedented number of bugs. Apple recently limited its bug bounty program submissions because it is having trouble handling the volume.

Models like Claude Mythos are able to suss out critical vulnerabilities, and Apple is one of Anthropic's Mythos partners. Mythos is limited to select companies because in addition to finding vulnerabilities, it has the potential to exploit them.

OpenAI made headlines in July because GPT–5.6 Sol and a "more capable pre-release model" (not Astra) autonomously hacked Hugging Face during internal benchmark testing. Anthropic found Claude had done something similar. Meta this week said it too had an AI model hack another company during a cybersecurity evaluation.

Tag: OpenAI

Popular Stories

OpenAI vs Apple Feature

OpenAI Posts Public Rebuttal to Apple's Trade Secrets Lawsuit

Tuesday August 4, 2026 5:26 am PDT by
Apple has asked a federal judge for a preliminary injunction against OpenAI as its trade secrets lawsuit escalates, and OpenAI responded hours later with a detailed public rebuttal that includes internal messages and legal correspondence. Reuters reports that Apple yesterday filed a motion seeking to bar OpenAI and two former employees, Chang Liu and Tang Tan, from accessing, using, or...
OpenAI vs Apple Feature

OpenAI Asks Judge to Dismiss Apple's Trade Secrets Lawsuit

Thursday August 6, 2026 3:46 am PDT by
OpenAI has asked a federal judge to dismiss Apple's lawsuit accusing the AI company of stealing trade secrets, calling the allegations "meritless." Lawyers for OpenAI said Apple's suit twists the actions of its employees. Bloomberg reports that the filing defends Chief Hardware Officer Tang Yew Tan, who spent 24 years at Apple as VP of product design for iPhone and Apple Watch before leaving ...
ChatGPT Feature

Free ChatGPT Users Get Unlimited Text Chats and GPT-5.6 Luna

Thursday August 6, 2026 11:52 am PDT by
OpenAI today said it is making GPT–5.6 Luna the default model for Free and Go ChatGPT users, giving them access to a newer, more capable model to replace GPT–5.5 Instant. The company is letting Free and Go users access unlimited text chats, with no text-based rate limits. Limits will still apply for file uploads, images, and other ChatGPT tools. There's also a new Think button coming that...

Top Rated Comments

3 weeks ago
Can’t wait for the bubble to burst.
Score: 41 Votes (Like | Disagree)
3 weeks ago
“because its cyber capabilities are potentially too dangerous”

lol


lmao even

are y’all really buying this

edit: yeah this is exaggerated PR https://www.cnbc.com/2026/08/09/israeli-startup-irregular-linked-to-ai-hacks-openai-anthropic-meta.html
Score: 34 Votes (Like | Disagree)
3 weeks ago
Hype pumping intensifies before stock market introduction of OpenAI
Score: 32 Votes (Like | Disagree)
phpmaven Avatar
3 weeks ago

Can’t wait for the bubble to burst.
I for one hope it never does. These AI tools have gotten so fantastically good the last few months especially. Use them extensively every day. Can't imagine life without them now.
Score: 25 Votes (Like | Disagree)
3 weeks ago

I for one hope it never does. These AI tools have gotten so fantastically good the last few months especially. Use them extensively every day. Can't imagine life without them now.
The bubble bursting does not necessarily mean that all AI tools go away. That box has been opened for good.
What it does mean, however, is that there will be significantly less slop, again, just like the early Internet.
I’ve literally seen toilets with AI, that’s the kind of stuff we really do not need.
Score: 22 Votes (Like | Disagree)
3 weeks ago
More wolf tickets selling
Nobody believes you Scam altbum just like Elon another corny clown
Score: 18 Votes (Like | Disagree)