Story · @OpenAI
OpenAI previews Astra safety evaluation as it reaches 'Critical' cyber threshold (@OpenAI)
X post · Story page
Ahead of releasing Astra, OpenAI says the model reaches the Critical threshold for cybersecurity under its Preparedness Framework and previews how it evaluated the model before shipping.
In plain words
- OpenAI says its coming Astra artificial intelligence system has made a major leap in computer security tasks.
- Before release, the company tested Astra using its Preparedness Framework, a process for measuring computer security abilities.
- Those tests placed Astra at the Critical level, the company's label for a serious degree of computer security ability.
- For future users, the checks are intended to support safer, broader access to increasingly capable artificial intelligence.
Appeared in
- Anthropic's deliberately misaligned model, Fable 5.1, and a fix for reward hacking
Sep 02, 2026 · from X
Subscribe
Get the brief in your inbox
Pick daily, weekly, or both. Nothing is gated either way: every issue is on the site and in the feeds.
- Weekdays at 8:45am IST, one lead story and 6 to 9 items.
- Sundays, an argued synthesis rather than a recap.
- One click to leave, and quiet days say so in the subject line.