OpenAI Details Astra Cybersecurity Evaluations and Safeguards
OpenAI shares preliminary evaluations for Astra and steps to strengthen safeguards against critical cyber capabilities.
OpenAI posted that it is sharing preliminary cybersecurity evaluations for its Astra model and the steps it is taking to strengthen safeguards. The company states it cannot rule out that Astra reaches critical cyber capabilities under its Preparedness Framework. It will impose expanded safety testing, isolated evaluation environments, and universal monitoring across agentic applications before any release. OpenAI says it is erring on the side of caution to develop Astra responsibly and share it with defenders.
After evaluating one of our upcoming models, Astra, we're treating it as our first "critical" model for cybersecurity under our Preparedness Framework. This is a scenario we've planned for, and we're putting additional controls in place to ensure Astra's further development…
