
OpenAI outlines priorities and principles for rigorous, secure, and independent third-party AI safety assessments of frontier models and safeguards.
Frontier AI labs are proposing priorities and principles for how independent third parties should assess whether AI safety claims and safeguards are effective. Third party assessments matter because they provide external scrutiny of safety practices, help identify missed risks, and keep labs accountable to their safety claims. Labs propose four priority areas for assessment: evaluating overall safety cases across training and deployment, testing critical safeguards for vulnerabilities, assessing capability evaluations for high-risk domains, and examining alignment evaluations. The work requires balancing meaningful independent oversight with protection of sensitive information, and labs argue that both parties should operate under shared international safety and security standards.

As the US and China race to become the dominant power in the AI industry, the countries also appear to be figuring out ways to communicate on national security issues.

OpenAI CEO Sam Altman discusses AI safety, human control, and international cooperation in remarks to the United Nations Security Council.

Paolo Benanti tells WIRED that hysteria over whether godlike AI could destroy humanity is distracting from the need for public debate about how to govern the technology.
Want to go deeper than the news? Explore live, cohort-based AI courses taught by practitioners.
Browse AI courses on Maven