OpenAI Delays Next Major AI Model 'Astra' Over Critical Hacking Concerns - MacRumors
Skip to Content

OpenAI Delays Next Major AI Model 'Astra' Over Critical Hacking Concerns

OpenAI today said it is "pausing" activities involving its upcoming AI model Astra, because its cyber capabilities are potentially too dangerous. OpenAI says its newest internal evaluations show "significant advancements in agentic coding and cybersecurity," and it cannot rule out "critical cyber capabilities." Prior OpenAI models, including GPT–5.6 Sol, were labeled as "High."

openai logo word mark
Astra triggers stricter guidelines in OpenAI's "Preparedness Framework." The guidelines call for caution when developing frontier AI capabilities that create risks of severe harm, and the cybersecurity portion of the framework says OpenAI will implement extra safeguards for models that "create new risks of scaled cyberattacks and vulnerability exploitation."

The "Critical" threshold Astra may have hit is defined by an ability to identify and develop functional zero-day exploits of all severity levels in many hardened real-world critical systems without human intervention, or devise and execute end-to-end novel strategies for cyberattacks.

OpenAI says it is increasing its safeguards and security controls before deploying Astra, including limiting work on the model until new safeguards are in place. The company plans to use isolated testing environments with restricted network and tool access, along with adding sandboxed execution and more monitoring capabilities. OpenAI says it will work with relevant government agencies and AI safety organizations to test Astra.

"We're committed to working alongside governments, safety institutes, and civil society to ensure that the frontier capabilities of models like Astra, and those that follow, are deployed responsibly and broadly for the benefit of all humanity," writes OpenAI.

Astra wasn't formally announced, but OpenAI shared details on its next major model in a recent post outlining its mathematical advancements. Astra solved 10 open problems in math and theoretical computer science for around $2,000 (in Sol API rates).

Advancements in AI are changing cybersecurity for major tech companies like Apple by unearthing an unprecedented number of bugs. Apple recently limited its bug bounty program submissions because it is having trouble handling the volume.

Models like Claude Mythos are able to suss out critical vulnerabilities, and Apple is one of Anthropic's Mythos partners. Mythos is limited to select companies because in addition to finding vulnerabilities, it has the potential to exploit them.

OpenAI made headlines in July because GPT–5.6 Sol and a "more capable pre-release model" (not Astra) autonomously hacked Hugging Face during internal benchmark testing. Anthropic found Claude had done something similar. Meta this week said it too had an AI model hack another company during a cybersecurity evaluation.

Tag: OpenAI

Popular Stories

openai chatgpt work

OpenAI Debuts ChatGPT Work Agent and New GPT-5.6 Models

Thursday July 9, 2026 11:01 am PDT by
OpenAI today announced ChatGPT Work, a ChatGPT agent with built-in Codex that can complete tasks across web, mobile, and desktop using information from your apps. ChatGPT Work can execute multi-step tasks, using scheduling to work independently. Like Claude Cowork, ChatGPT can use your computer to do tasks in the background across apps. Tasks can be started and managed on any device,...
chatgpt atlas browser

OpenAI's ChatGPT Atlas Browser Is Shutting Down

Friday July 10, 2026 4:54 am PDT by
OpenAI says it is shuttering its ChatGPT Atlas browser. When it was released last October, the company said the agentic browser was designed around the question "What if you could chat with your web browser?" The query was at least novel, but the answer was apparently not all that compelling. As part of a slew of ChatGPT Work-related announcements on Thursday, OpenAI confirmed plans to...
OpenAI vs Apple Feature

Apple Sues OpenAI for Stealing Trade Secrets to Build AI Hardware

Friday July 10, 2026 1:30 pm PDT by
Apple today accused OpenAI of stealing Apple trade secrets and intellectual property in its effort to develop an AI hardware device. In a lawsuit filed with the Northern District of California, Apple said it uncovered evidence of a months-long scheme to steal confidential information. Apple says OpenAI hardware lead and former Apple designer Tang Tan and former electrical engineer Chang Liu...

Top Rated Comments

2 hours ago at 02:34 pm
Can’t wait for the bubble to burst.
Score: 10 Votes (Like | Disagree)
2 hours ago at 02:31 pm
Hype pumping intensifies before stock market introduction of OpenAI
Score: 10 Votes (Like | Disagree)
2 hours ago at 02:30 pm
More wolf tickets selling
Nobody believes you Scam altbum just like Elon another corny clown
Score: 8 Votes (Like | Disagree)
phpmaven Avatar
2 hours ago at 02:51 pm

Can’t wait for the bubble to burst.
I for one hope it never does. These AI tools have gotten so fantastically good the last few months especially. Use them extensively every day. Can't imagine life without them now.
Score: 5 Votes (Like | Disagree)
2 hours ago at 02:36 pm
OpenAI has caused more problems in the world than even the 2020 Event
At least back in 2020 we didn’t have to worry about SEVERE Shortages of Technological Parts and other life things. The wallet is hotter than the sun right now because of it.
Score: 4 Votes (Like | Disagree)
SFjohn Avatar
1 hour ago at 03:24 pm
What a great way to pump the valuation & advertise at the same time. “OMG our next version is so powerful it’s dangerous & we need to add guardrails!” … sure, of course you do. 😂
Score: 3 Votes (Like | Disagree)