Provision record
Anthropic · Anthropic Responsible Scaling Policy · View original document ↗

Misalignment Risk Affirmative Case Requirement at AI R&D-4

High severity High confidence Explicit document language Common · 217 of 352 platforms

Key Facts

When must Anthropic develop an affirmative case identifying misalignment risks?
Anthropic requires that, once models cross the AI R&D-4 capability threshold, it develops an affirmative case identifying the most immediate and relevant misalignment risks from models pursuing misaligned goals and explaining how those risks have been mitigated.
What must Anthropic develop once models cross the AI R&D-4 capability threshold?
Anthropic requires that, once models cross the AI R&D-4 capability threshold, it develops an affirmative case identifying the most immediate and relevant misalignment risks from models pursuing misaligned goals and explaining how those risks have been mitigated.
Stay ahead of the changes
Track Anthropic and get the diff the day its terms change.
Share 𝕏 Share in Share 🔒 PDF

This analysis describes what Anthropic's agreement states, permits, or reserves. It does not constitute a legal determination about enforceability. Regulatory applicability and practical outcomes may vary by jurisdiction, enforcement context, and individual circumstances. Read our methodology

ConductAtlas Analysis

Why it matters (compliance & governance perspective)

The requirement creates a formal, documented obligation for Anthropic to affirmatively demonstrate misalignment risk awareness and mitigation at a specific capability level, rather than leaving it to discretionary assessment.

Consumer impact (what this means for users)

Anthropic is required to produce a documented affirmative case on misalignment risks and mitigation once its models reach the AI R&D-4 capability threshold.

How other platforms handle this

Tinder Medium

This is still Your Content, and you are responsible for it and its accuracy, as well as your use of it on our Services and any and all decisions made, actions taken, and failures to take action based on Your Content.

ActiveCampaign Medium

Due to the nature of the AI Features, generated Marketing Content may not be unique across users and the AI Features may generate the same or similar Marketing Content for other users.

GitHub Medium

We generate new information from other data we collect to derive likely preferences or other characteristics. For instance, we infer your general geographic location based on your IP address.

See all platforms with this clause type →
▸ View Original Clause Language DOCUMENT RECORD
"
once models cross the AI R&D-4 capability threshold, we develop an affirmative case identifying the most immediate and relevant misalignment risks from models pursuing misaligned goals and explaining how we have mitigated them

Excerpt from Anthropic's Responsible Scaling Policy

Applicable regulations

EU AI Act
European Union
California AB 2013 AI Training Data Transparency
US-CA
Colorado AI Act
US-CO
EU AI Act - High Risk Provisions
EU
GDPR
European Union
Texas AI Act
Texas, USA
Trump Executive Order on AI Policy Framework
US
UK GDPR
United Kingdom

Provision details

Document information
Document
Anthropic Responsible Scaling Policy
Entity
Anthropic
Document last updated
May 12, 2026
Tracking information
First tracked
July 12, 2026
Last verified
July 12, 2026
Record ID
CA-P-073920
Document ID
CA-D-00823
Evidence Provenance
Source URL
Wayback Machine
Content hash (SHA-256)
9ad6c66902eb2b771ccc40a9d0e1c2d4664c197372e6833be51b902e16754d55
Analysis generated
July 12, 2026 14:38 UTC
Methodology
Evidence
✓ Snapshot stored   ✓ Hash verified
Citation Record
Entity: Anthropic
Document: Anthropic Responsible Scaling Policy
Record ID: CA-P-073920
Captured: 2026-07-12 14:38:45 UTC
SHA-256: 9ad6c66902eb2b77…
URL: https://conductatlas.com/platform/anthropic/anthropic-responsible-scaling-policy/provision/CA-P-073920/misalignment-risk-affirmative-case-requirement-at-ai-rd-4/
Accessed: Aug. 18, 2026
Permanent archival reference. Stable identifier suitable for legal filings, compliance documentation, and research citation.
Classification
Severity
High
Categories

Other risks in this policy

Related Analysis

Get the research letter

Companies change their terms quietly. We read every version and catch what actually changed. One email a week on the changes that matter and what they mean.

Frequently Asked Questions

What does Anthropic's Misalignment Risk Affirmative Case Requirement at AI R&D-4 clause do?

The requirement creates a formal, documented obligation for Anthropic to affirmatively demonstrate misalignment risk awareness and mitigation at a specific capability level, rather than leaving it to discretionary assessment.

How does this clause affect you?

Anthropic is required to produce a documented affirmative case on misalignment risks and mitigation once its models reach the AI R&D-4 capability threshold.

How many platforms have this type of clause?

ConductAtlas has identified this type of provision across 217 platforms. See the full comparison.

Is ConductAtlas affiliated with Anthropic?

No. ConductAtlas is an independent monitoring service. We are not affiliated with, endorsed by, or sponsored by Anthropic.