Key takeaways

  • Cybersecurity is rapidly changing as models become more capable in ways that can both strengthen cyberdefenses and enable attacks at…
  • These results, in addition to expert assessments, have led us to conclude last night that we cannot rule out critical cyber capabilities…
  • We first published our Preparedness Framework in December 2023, well before models approached biological, chemical, cybersecurity, and AI…

What happened

Cybersecurity is rapidly changing as models become more capable in ways that can both strengthen cyberdefenses and enable attacks at unprecedented speed and scale. Our latest internal evaluations of Astra, one of our upcoming models, over the past few days indicate significant advancements in agentic coding and cybersecurity.

While we continue to benchmark and assess this model, our preliminary evaluations indicate strong enough performance that we cannot rule out Critical capability level at this time. Astra is an upcoming model, and was not involved in exploiting Hugging Face. Accordingly, we have scaled up robustness testing of our safeguards and security controls so that they are appropriate for a deployment of these capabilities.

Why it matters

These results, in addition to expert assessments, have led us to conclude last night that we cannot rule out critical cyber capabilities under our Preparedness Framework⁠(opens in a new window). We are sharing this because we believe it’s important to be transparent with the public and the safety and security communities about this potential shift in capabilities.

We first published our Preparedness Framework in December 2023, well before models approached biological, chemical, cybersecurity, and AI self-improvement capabilities at this level. We created it to give us a guide for identifying progress in capability and then planning what our company would do as those capabilities emerge. 6‑Sol, have been evaluated for frontier cyber capabilities and assessed at the High (rather than Critical) threshold.

Under our Preparedness Framework, a model reaches the Critical cybersecurity threshold if it can identify and develop functional zero-day exploits of all severity levels in many hardened real-world critical systems without human intervention, or can devise and execute end-to-end novel strategies for cyberattacks against hardened targets given only a high level desired goal.

What to watch

Internally, we have also taken the following steps so that further development of this model happens safely and securely: The framework has already guided us through other capability transitions. In June 2025, as our models approached the high capability threshold for biology under the Preparedness Framework, we outlined the steps we were taking to strengthen safeguards, expand testing, work with external experts, and deploy additional security controls.

We are applying the same principle here. We believe advanced cyber-capable models should help defenders identify and address vulnerabilities before attackers do. We’re committed to working alongside governments, safety institutes, and civil society to ensure that the frontier capabilities of models like Astra, and those that follow, are deployed responsibly and broadly for the benefit of all humanity.