tech
December 16, 2025
Announcing our updated Responsible Scaling Policy
Today we are publishing a significant update to our Responsible Scaling Policy (RSP), the risk governance framework we use to mitigate potential catastrophic risks from frontier AI systems. This update introduces a more flexible and nuanced approach to assessing and managing AI risks while maintaining our commitment not to train or deploy models unless we have implemented adequate safeguards. Key improvements include new capability thresholds to indicate when we will upgrade our safeguards, refined processes for evaluating model capabilities and the adequacy of our safeguards (inspired by safety case methodologies), and new measures for internal governance and external input. By learning from our implementation experiences and drawing on risk management practices used in other high-consequence industries, we aim to better prepare for the rapid pace of AI advancement.

TL;DR
- Anthropic updated its Responsible Scaling Policy (RSP) for managing risks from frontier AI systems.
- The RSP now includes new capability thresholds to determine when safeguards need to be upgraded.
- Processes for evaluating model capabilities and safeguard adequacy have been refined, inspired by safety case methodologies.
- Key capability thresholds identified are for Autonomous AI Research and Development and assistance in creating CBRN weapons.
- The company has established processes for capability and safeguard assessments, along with measures for internal governance and external input.
- Lessons learned from the previous year's implementation highlighted the need for policy flexibility and improved compliance tracking.
- Jared Kaplan will serve as Anthropic's Responsible Scaling Officer, succeeding Sam McCandlish.
Continue reading
the original article