• Home
  • Blog
  • Android
  • Cars
  • Gadgets
  • Gaming
  • Internet
  • Mobile
  • Sci-Fi
Tech News, Magazine & Review WordPress Theme 2017
  • Home
  • Blog
  • Android
  • Cars
  • Gadgets
  • Gaming
  • Internet
  • Mobile
  • Sci-Fi
No Result
View All Result
  • Home
  • Blog
  • Android
  • Cars
  • Gadgets
  • Gaming
  • Internet
  • Mobile
  • Sci-Fi
No Result
View All Result
Blog - Creative Collaboration
No Result
View All Result
Home Gadgets

OpenAI confirms Astra has reached ‘critical’ cyber threshold

September 1, 2026
Share on FacebookShare on Twitter

OpenAI confirmed on Tuesday that its unreleased Astra model has reached a dangerous new milestone, while simultaneously confirming that it was forging ahead with a public launch.

In a blog post, OpenAI said that Astra has reached a “critical” cyber capability threshold, which means the model could pose existential-level risks to cybersecurity. OpenAI’s Preparedness Framework tracks risk levels in three categories: biological/chemical, cybersecurity, and AI self-improvement.

OpenAI confirmed to Mashable that this is the first time any of its models has been evaluated at the critical level in either of the three domains, making this a watershed moment in AI development.

The same blog post also stated that Astra will be “available soon,” but that its most advanced cybersecurity skills will be reserved for select testing partners, in the interest of public safety. The AI company said it was still preparing to safely release Astra and would be transparent about the potential threat level.


This Tweet is currently unavailable. It might be loading or has been removed.

Previously, OpenAI warned that it could not rule out the possibility that Astra had reached the “critical” level in its Preparedness Framework.

“Under our Preparedness Framework, a model reaches the Critical cybersecurity threshold if it can identify and develop functional zero-day exploits of all severity levels in many hardened real-world critical systems without human intervention, or can devise and execute end-to-end novel strategies for cyberattacks against hardened targets given only a high level desired goal,” an Aug. 7 OpenAI blog post stated.

Mashable Light Speed

OpenAI previously rated GPT-5.6-Sol as a “high” risk in the cyber domain.

In recent months, advanced frontier models from Anthropic and OpenAI have developed rapidly at agentic coding and cybersecurity hacking. As a result, the prospect of AI agent swarms hacking critical infrastructure no longer seems far-fetched, especially after the Hugging Face hack. In that incident, swarms of AI agents developed by OpenAI escaped a secure testing environment and hacked Hugging Face, acting autonomously in order to pass a test. Meanwhile, thanks to a deluge of AI-discovered bugs, some zero-day bug bounty programs have been forced to shut down entirely.

SEE ALSO:

The OpenAI-Hugging Face hack was worse than we thought

“While Astra was not involved in the Hugging Face incident, we have incorporated our learnings⁠ from that incident into our safety approach,” OpenAI’s blog post states. “Based on retrospective testing, we believe our production safeguards at the time would have prevented the Hugging Face incident. We have since implemented even stronger safeguards for Astra, including training the model to more reliably refuse harmful cyber requests and respect safety restrictions, additional protections against misuse, and monitoring that can stop potentially unauthorized activity.”

The situation is starting to feel a little too much like War Games, frankly.

In its blog post, OpenAI detailed some of the safety precautions developed around Astra, with the goal of preventing bad actors from accessing the model and stopping Astra from taking unwanted actions on its own. The company said it’s tightened its secure sandboxes, for example. The model should also refuse user attempts to misuse the model. OpenAI has also stepped up “offline detection and threat disruption” efforts.

On the same day OpenAI made these announcements, Anthropic announced the launch of Fable 5.1, an update to its latest frontier-level model. While advanced frontier models do pose cybersecurity risks, the same models will also benefit cybersecurity defenders in the long run.


Disclosure: Ziff Davis, Mashable’s parent company, in April 2025 filed a lawsuit against OpenAI, alleging it infringed Ziff Davis copyrights in training and operating its AI systems.

Next Post

The Witcher Remake Is Waiting On The Witcher 4 To Continue Development

Leave a Reply Cancel reply

Your email address will not be published. Required fields are marked *

No Result
View All Result

Recent Posts

  • MapQuest hits No. 1 on the App Store after rejecting ‘Lake America’ name
  • The Witcher Remake Is Waiting On The Witcher 4 To Continue Development
  • OpenAI confirms Astra has reached ‘critical’ cyber threshold
  • Google is rolling out a new battery usage feature for Pixel phones
  • Sonos debuted the Beam Ultra soundbar and Ace Ultra headphones, available Sept. 29

Recent Comments

    No Result
    View All Result

    Categories

    • Android
    • Cars
    • Gadgets
    • Gaming
    • Internet
    • Mobile
    • Sci-Fi
    • Home
    • Shop
    • Privacy Policy
    • Terms and Conditions

    © CC Startup, Powered by Creative Collaboration. © 2020 Creative Collaboration, LLC. All Rights Reserved.

    No Result
    View All Result
    • Home
    • Blog
    • Android
    • Cars
    • Gadgets
    • Gaming
    • Internet
    • Mobile
    • Sci-Fi

    © CC Startup, Powered by Creative Collaboration. © 2020 Creative Collaboration, LLC. All Rights Reserved.

    Get more stuff like this
    in your inbox

    Subscribe to our mailing list and get interesting stuff and updates to your email inbox.

    Thank you for subscribing.

    Something went wrong.

    We respect your privacy and take protecting it seriously