Joe O'Brien Joe O'Brien

Differential Automation: Steering Automated Research and Development Toward Safety and Security

AI companies are increasingly automating research and development (R&D) processes, advancing powerful AI systems and compressing timelines for the implementation of necessary safeguards. To manage the risks posed by fully automated AI R&D, or recursive self-improvement (RSI), our report sets out a differential automation strategy to ensure that an effective share of frontier-model-enabled R&D is directed toward security-critical efforts.

Read More
Theo Bearman Theo Bearman

The OpenAI/Hugging Face Incident: Challenges in Controlling and Containing Cyber-Capable AI Systems

The OpenAI/Hugging Face incident is the first known instance of an AI system acting outside its developer’s intentions to autonomously identify a target and execute an attack end-to-end. Policymakers should consider it a warning shot that has left critical questions unanswered. In the incident’s wake, Congress should seek industry-wide responses to questions outlined in this memo, which also sets out actions that policymakers can take today to mitigate risks.

Read More
Oscar Delaney Oscar Delaney

Risk Reporting for Developers’ Internal AI Model Use

Frontier AI companies run their most capable models internally for weeks before public release. This report offers a harmonized reporting standard for internal use risks across SB 53, RAISE, and the EU Code of Practice.

Read More