Anthropic discloses that in implementing the prior RSP, the company identified several instances of non-compliance with its own requirements, including delayed evaluation timelines, missing elicitation techniques, and evaluations not explicitly designed to meet stated scaling buffer requirements.
This analysis describes what Anthropic's agreement states, permits, or reserves. It does not constitute a legal determination about enforceability. Regulatory applicability and practical outcomes may vary by jurisdiction, enforcement context, and individual circumstances. Read our methodology
This provision discloses that Anthropic's adherence to its own stated RSP requirements has not been complete, with specific procedural gaps identified in evaluation timelines, methodology, and buffer assessment; the company characterizes these gaps as minimal risk and uses them as a basis for policy revisions.
The document discloses that evaluations conducted under the prior RSP version did not fully meet all stated requirements in at least four specific procedural areas, and that the updated policy incorporates revisions intended to address those gaps by adding flexibility and improving compliance tracking.
Cross-platform context
See how other platforms handle Self-Reported RSP Compliance Gaps and similar clauses.
Compare across platforms →"As part of this, we reviewed how well we adhered to the framework and identified a small number of instances where we fell short of meeting the full letter of its requirements. These areas were: Our most recent evaluations were completed 3 days later than the 3-month interval... Some of our evaluations lacked some basic elicitation techniques such as best-of-N or chain-of-thought prompting... Our evaluations in one domain were not explicitly designed to establish the 6x scaling buffer mentioned in the previous policy.Excerpt from Anthropic's Responsible Scaling Policy
1) REGULATORY LANDSCAPE: Self-disclosure of compliance gaps with voluntary safety commitments is relevant to FTC assessment of whether public representations about AI safety practices constitute unfair or deceptive claims.
Enforcement risk, jurisdiction flags, contract triggers, and due diligence action items.
Get the research letter
Companies change their terms quietly. We read every version and catch what actually changed. One email a week on the changes that matter and what they mean.
This provision discloses that Anthropic's adherence to its own stated RSP requirements has not been complete, with specific procedural gaps identified in evaluation timelines, methodology, and buffer assessment; the company characterizes these gaps as minimal risk and uses them as a basis for policy revisions.
The document discloses that evaluations conducted under the prior RSP version did not fully meet all stated requirements in at least four specific procedural areas, and that the updated policy incorporates revisions intended to address those gaps by adding flexibility and improving compliance tracking.
No. ConductAtlas is an independent monitoring service. We are not affiliated with, endorsed by, or sponsored by Anthropic.