Provision record
Anthropic · Claude Opus 4.8 System Card · View original document ↗

CB-1 Biological Capability Threshold and Classifier Mitigation Reliance

High severity High confidence Explicit document language Unique · 0 of 352 platforms
Stay ahead of the changes
Track Anthropic and get the diff the day its terms change.
Share 𝕏 Share in Share 🔒 PDF
Document Record

What it is

The document states that Opus 4.8 meets the CB-1 capability threshold, meaning it can provide specific and actionable information relevant to biological weapons production, and that Anthropic's primary mitigation relies on real-time classifier guards, access controls, a bug bounty program, and rapid response options for jailbreaks rather than capability removal.

ⓘ

This analysis describes what Anthropic's agreement states, permits, or reserves. It does not constitute a legal determination about enforceability. Regulatory applicability and practical outcomes may vary by jurisdiction, enforcement context, and individual circumstances. Read our methodology

ConductAtlas Analysis

Why it matters (compliance & governance perspective)

This provision establishes that a deployed production model has been assessed as capable of materially assisting biological weapons-relevant tasks, with risk management delegated to runtime detection and blocking mechanisms. The document explicitly acknowledges that catastrophic risk under this framework is 'very low but not negligible,' which represents a material disclosure for institutional deployers and regulatory bodies evaluating the adequacy of classifier-based mitigations relative to capability-level controls.

Consumer impact (what this means for users)

Under these terms, users interacting with Opus 4.8 in biological sciences contexts may encounter classifier-triggered refusals or access restrictions without receiving explicit explanation of the underlying risk classification. The agreement reserves Anthropic's right to apply and adjust these classifier guards without advance notice to users.

Cross-platform context

See how other platforms handle CB-1 Biological Capability Threshold and Classifier Mitigation Reliance and similar clauses.

Compare across platforms →
▸ View Original Clause Language DOCUMENT RECORD
"
Our capability assessments are consistent with the model being capable of providing specific, actionable information relevant to the threat model, such that it may save even experts in these domains substantial time. As with other models with these properties, we apply strong real-time classifier guards to this model and access controls for classifier guard exemptions. We also maintain a bug bounty program and threat intelligence for continual assessment of our classifier guards' effectiveness; a variety of rapid response options for jailbreaks; and security controls to reduce risk of model weight theft. We believe these risk mitigations are equal to or stronger than our historical ASL-3 protections and sufficient to make catastrophic risk in this category very low but not negligible.

Excerpt from Anthropic's Claude Opus 4.8 System Card

ConductAtlas Analysis

Institutional analysis (regulatory & governance intelligence)

(1) REGULATORY LANDSCAPE: This provision is materially relevant to the EU AI Act's systemic risk assessment requirements for general-purpose AI models, which may require documented capability evaluations and mitigation measures for models with identified high-risk …

Insight

Unlock the full institutional analysis

Enforcement risk, jurisdiction flags, contract triggers, and due diligence action items.

Applicable agencies

  • Federal Trade Commission (ftc)
    Oversees unfair or deceptive business practices and can investigate companies that mislead consumers about data collection, sharing, or use.
    Who can file: Anyone affected by the company's practices (US or international)
    What you need: Your account details, a timeline of relevant events, and a description of the specific issue
    What to expect: Complaints inform FTC enforcement priorities and investigations but do not result in individual resolution or compensation
    File a complaint →

Provision details

Document information
Document
Claude Opus 4.8 System Card
Entity
Anthropic
Document last updated
July 6, 2026
Tracking information
First tracked
July 7, 2026
Last verified
July 7, 2026
Record ID
CA-P-013476
Document ID
CA-D-00920
Evidence Provenance
Source URL
Wayback Machine
Content hash (SHA-256)
7f7ede707e6d4291941e66235c56ac7efc7262eaa7866f61c3a4e354180f417f
Analysis generated
July 7, 2026 23:44 UTC
Methodology
Evidence
✓ Snapshot stored   ✓ Hash verified
Citation Record
Entity: Anthropic
Document: Claude Opus 4.8 System Card
Record ID: CA-P-013476
Captured: 2026-07-07 23:44:13 UTC
SHA-256: 7f7ede707e6d4291…
URL: https://conductatlas.com/platform/anthropic/claude-opus-48-system-card/provision/CA-P-013476/cb-1-biological-capability-threshold-and-classifier-mitigation-reliance/
Accessed: Oct. 2, 2026
Permanent archival reference. Stable identifier suitable for legal filings, compliance documentation, and research citation.
Classification
Severity
High
Categories

Other risks in this policy

Get the research letter

Companies change their terms quietly. We read every version and catch what actually changed. One email a week on the changes that matter and what they mean.

Frequently Asked Questions

What does Anthropic's CB-1 Biological Capability Threshold and Classifier Mitigation Reliance clause do?

This provision establishes that a deployed production model has been assessed as capable of materially assisting biological weapons-relevant tasks, with risk management delegated to runtime detection and blocking mechanisms. The document explicitly acknowledges that catastrophic risk under this framework is 'very low but not negligible,' which represents a material disclosure for institutional deployers and regulatory bodies evaluating the adequacy of …

How does this clause affect you?

Under these terms, users interacting with Opus 4.8 in biological sciences contexts may encounter classifier-triggered refusals or access restrictions without receiving explicit explanation of the underlying risk classification. The agreement reserves Anthropic's right to apply and adjust these classifier guards without advance notice to users.

Is ConductAtlas affiliated with Anthropic?

No. ConductAtlas is an independent monitoring service. We are not affiliated with, endorsed by, or sponsored by Anthropic.