OpenAI Details Astra Cybersecurity Evaluations and Safeguards
OpenAI shares preliminary evaluations for Astra and steps to strengthen safeguards against critical cyber capabilities.
TLDR
OpenAI posted that it is sharing preliminary cybersecurity evaluations for its Astra model and the steps it is taking to strengthen safeguards. The company states it cannot rule out that Astra reaches critical cyber capabilities under its Preparedness Framework. It will impose expanded safety testing, isolated evaluation environments, and universal monitoring across agentic applications before any release. OpenAI says it is erring on the side of caution to develop Astra responsibly and share it with defenders.