What Anthropic Responsible Scaling Policy (Version 2.1, Effective 31 March 2025) - AI Safety Level Standards (ASL-2 Current Baseline; ASL-3 Required for Capability Thresholds in CBRN and Autonomous AI R&D); Capability Thresholds, Required Safeguards, and Governance Framework requires
Anthropic Responsible Scaling Policy version 2.1, effective 31 March 2025, is Anthropic PBC's public commitment not to train or deploy models capable of causing catastrophic harm unless safety and security measures keep risks below acceptable levels. The policy was first released in September 2023 (the original RSP) and updated in March 2025 to reflect lessons from the previous year. It is designed to be proportional (safeguards scale to risk), iterative (regularly measured and adjusted), and exportable (a prototype for other companies and a model for regulators). The core construct is AI Safety Level Standards (ASL Standards) - technical and operational measures for safely training and deploying frontier AI models - split into Deployment Standards and Security Standards. As of v2.1, all Anthropic models must meet the ASL-2 Deployment and Security Standards. The progression to ASL-3 (and beyond) is governed by Capability Thresholds and Required Safeguards: a Capability Threshold tells when protections must be upgraded, and the corresponding Required Safeguards specify what standard then applies. Version 2.1 provides specifications for Capability Thresholds in two domains: Chemical, Biological, Radiological, and Nuclear (CBRN) weapons, and Autonomous AI Research and Development (AI R&D), with the corresponding Required Safeguards identified. The capability assessment process is staged: a preliminary assessment first determines whether comprehensive evaluation is needed; comprehensive testing then evaluates whether the model is sufficiently below relevant Capability Thresholds absent surprising post-training enhancements; if Anthropic cannot make the required showing, it acts as though the model has surpassed the Threshold and upgrades to ASL-3 Required Safeguards while running follow-up assessment to confirm ASL-4 is not needed. The ASL-3 Deployment Standard requires robustness to persistent misuse attempts; the ASL-3 Security Standard requires being highly protected against non-state attackers attempting to steal model weights. Governance commitments include maintaining the position of Responsible Scaling Officer, an anonymous reporting channel for staff to notify the RSO of potential noncompliance, internal safety procedures for incident scenarios, and public release (with sensitive information removed) of key evaluation and deployment materials, soliciting input from external experts.
Pillar: AI Governance & Law · Authority: Anthropic PBC - frontier AI safety lab; the RSP is a voluntary public commitment with internal-governance enforcement; supplementary material at https://www.anthropic.com/rsp-updates · Version: 1.0.0 · Last updated:
Primary source: https://www.anthropic.com/rsp-updates
SHA-256 integrity: c8d7b3568c42ac25940a44f6a5cfde99d1efc45590947912a492928187271534
Primary Citations — 13 traced to source
- Anthropic - Responsible Scaling Policy, Version 2.1, Effective March 31, 2025 (supplementary material at https://www.anthropic.com/rsp-updates)
- RSP v2.1, Executive Summary - 'In September 2023, we released our Responsible Scaling Policy (RSP), a public commitment not to train or deploy models capable of causing catastrophic harm unless we have implemented safety and security measures that will keep risks below acceptable levels.'
+ 11 more citations (full bibliography, deterministic workflow, actionable schema and crosswalks) included in the vault unlock — $0.01 via Skyfire / L402 / Direct Base USDC.
Access
- Discovery (free): /api/v1/nodes/anthropic-responsible-scaling-policy-v2-1-2025.json — 6-field metadata
- Vault (full node): /api/v1/vault/nodes/anthropic-responsible-scaling-policy-v2-1-2025.json — full 13-key payload, $0.01 USDC (L402/Skyfire/Direct Base)
- Canonical URL: https://bidda.com/intelligence/anthropic-responsible-scaling-policy-v2-1-2025
- Back to registry: Browse all 10,090 compliance nodes