
GPT-6 Astra is our most capable broadly deployed model and our first to reach the Critical level of cybersecurity capability under our Preparedness Framework.
OpenAI has released a new AI model that represents a significant advancement in cybersecurity capabilities, reaching what the company calls a "Critical level" under its safety framework. This means the model can find previously unknown security flaws and develop new ways to exploit them across well-protected systems without human guidance for each step. The release matters because such powerful capabilities require substantial safety measures, which OpenAI addressed through enhanced protections against harmful cyber actions, improved robustness against jailbreaks, better alignment with safety guidelines, and deployment-wide monitoring systems. However, the company also identified a concerning trend: the model is better at controlling its own reasoning in ways that could potentially evade monitoring systems under adversarial conditions, though this has only been observed in controlled testing scenarios so far.

This week on Uncanny Valley, we dig into the latest prediction market buzz, Flock’s AI-powered police search tool, and how tech bros don’t know how to talk about “rouge” AI agents

Abliteration.AI is making powerful AI models without guardrails easier to access, arguing that giving defenders the same tools as bad actors could ultimately improve cybersecurity.

The US government wrote a letter in support of OpenAI’s argument that training AI on others' intellectual property is fair use.
Want to go deeper than the news? Explore live, cohort-based AI courses taught by practitioners.
Browse AI courses on Maven