Anthropic · Claude Opus 4.8 System Card · View original document ↗

Agentic Safety and Prompt Injection Vulnerability Disclosure

High severity High confidence Explicitdocumentlanguage Unique · 0 of 352 platforms
Get alerted the next time Anthropic changes these terms. Get same-day alerts →
Share 𝕏 Share in Share 🔒 PDF
Recent governance activity Anthropic recorded 3 documented changes in the last 30 days.
Get same-day alerts →
Monitor governance changes for Anthropic Monitor emails you the same day this changes. The archive stays free.
Get same-day alerts →

Get the weekly research letter

Companies change their terms quietly. We read every version and catch what actually changed. One email a week on the changes that matter and what they mean. No account.

Document Record

What it is

The document states that Opus 4.8 exhibits weaker prompt injection robustness than its predecessor Opus 4.7 in some agentic contexts, and that applied safeguards partially close this gap. Anthropic conducted a one-week live bug bounty competition to assess robustness against adaptive attackers across coding, computer use, and browser use surfaces.

This analysis describes what Anthropic's agreement states, permits, or reserves. It does not constitute a legal determination about enforceability. Regulatory applicability and practical outcomes may vary by jurisdiction, enforcement context, and individual circumstances. Read our methodology

ConductAtlas Analysis

Why it matters (compliance & governance perspective)

This provision establishes that a production model deployed in agentic contexts has disclosed prompt injection vulnerabilities relative to its predecessor, with risk mitigation relying on applied safeguards rather than capability-level robustness improvements. Organizations deploying Opus 4.8 in autonomous or semi-autonomous agentic pipelines should assess this disclosure in the context of their specific deployment architectures.

Consumer impact (what this means for users)

Under these terms, users and developers deploying Opus 4.8 in agentic contexts, including Claude Code and computer use applications, operate under disclosed prompt injection risk that the document states is partially but not fully mitigated by applied safeguards. The bug bounty program provides a mechanism for security researchers to report discovered vulnerabilities.

Cross-platform context

See how other platforms handle Agentic Safety and Prompt Injection Vulnerability Disclosure and similar clauses.

Compare across platforms →

Monitoring

Anthropic has changed this document before.

Receive same-day alerts, structured change summaries, and monitoring for up to 25 platforms.

Get Monitor Or create a free account →
▸ View Original Clause Language DOCUMENT RECORD
"
Although it shows improvements in some areas (such as refusing malicious requests), we found Opus 4.8 to be somewhat less robust than Opus 4.7 in several agentic contexts (such as vulnerability to prompt injection attacks). However, the application of our safeguards closes the gap between the models in practice. We report the results of our first one-week live bug bounty for prompt injection.

Excerpt from Anthropic's Claude Opus 4.8 System Card

ConductAtlas Analysis

Institutional analysis (regulatory & governance intelligence)

(1) REGULATORY LANDSCAPE: Prompt injection vulnerabilities in agentic AI systems are relevant to emerging regulatory guidance on AI system security, including the EU AI Act's security and robustness requirements for high-risk AI systems, and NIST AI Risk Management Framework recommendations. The FTC's unfair or deceptive practices authority may also engage disclosures about known security vulnerabilities in deployed AI systems. (2) GOVERNANCE EXPOSURE: High. The document discloses that a production model in agentic deployment contexts has lower prompt injection robustness than its predecessor, with safeguards rather than model-level improvements as the primary mitigation. Organizations deploying Opus 4.8 in agentic workflows with access to sensitive data, financial systems, or operational infrastructure face heightened exposure to prompt injection-based misuse. (3) JURISDICTION FLAGS: EU/EEA deployments under the AI Act's security requirements for high-risk AI systems, and US deployments in regulated sectors including finance and healthcare, create heightened exposure for agentic applications subject to prompt injection risk. The one-week bug bounty duration may be assessed by external reviewers as a limited timeframe for comprehensive vulnerability discovery. (4) CONTRACT AND VENDOR IMPLICATIONS: B2B customers deploying Opus 4.8 in agentic pipelines should assess whether their service agreements with Anthropic address liability allocation for losses arising from prompt injection attacks, and whether the disclosed vulnerability affects their own customer-facing security representations. Vendors building on Claude Code or computer use APIs should review their own terms of service in light of this disclosure. (5) COMPLIANCE CONSIDERATIONS: Compliance teams should assess whether the agentic safety disclosure requires updating their internal AI risk registers, incident response plans, or vendor security assessments. Organizations in regulated sectors should evaluate whether prompt injection vulnerability in deployed agentic systems requires notification to regulators or customers under applicable security incident disclosure frameworks.

Full institutional analysis
Regulatory citations, enforcement risk, and due diligence action items.
Start Professional · $99/mo Start with Monitor · $29/mo

Applicable agencies

  • FTC
    The FTC's authority over unfair or deceptive practices and data security may engage disclosures about known prompt injection vulnerabilities in deployed AI systems, particularly where those systems process consumer data.
    File a complaint →

Provision details

Document information
Document
Claude Opus 4.8 System Card
Entity
Anthropic
Document last updated
July 6, 2026
Tracking information
First tracked
July 7, 2026
Last verified
July 7, 2026
Record ID
CA-P-013479
Document ID
CA-D-00920
Evidence Provenance
Source URL
Wayback Machine
Content hash (SHA-256)
7f7ede707e6d4291941e66235c56ac7efc7262eaa7866f61c3a4e354180f417f
Analysis generated
July 7, 2026 23:44 UTC
Methodology
Evidence
✓ Snapshot stored   ✓ Hash verified
Citation Record
Entity: Anthropic
Document: Claude Opus 4.8 System Card
Record ID: CA-P-013479
Captured: 2026-07-07 23:44:13 UTC
SHA-256: 7f7ede707e6d4291…
URL: https://conductatlas.com/platform/anthropic/claude-opus-48-system-card/provision/CA-P-013479/agentic-safety-and-prompt-injection-vulnerability-disclosure/
Accessed: July 23, 2026
Permanent archival reference. Stable identifier suitable for legal filings, compliance documentation, and research citation.
Classification
Severity
High
Categories

Other risks in this policy

Governance intelligence across arbitration, AI governance, data rights, indemnification, and retention
Provision-level monitoring, governance timelines, and regulatory mapping built from archived source documents and historical version tracking.
Start Professional · $99/mo Start with Monitor · $29/mo

Frequently Asked Questions

What does Anthropic's Agentic Safety and Prompt Injection Vulnerability Disclosure clause do?

This provision establishes that a production model deployed in agentic contexts has disclosed prompt injection vulnerabilities relative to its predecessor, with risk mitigation relying on applied safeguards rather than capability-level robustness improvements. Organizations deploying Opus 4.8 in autonomous or semi-autonomous agentic pipelines should assess this disclosure in the context of their specific deployment architectures.

How does this clause affect you?

Under these terms, users and developers deploying Opus 4.8 in agentic contexts, including Claude Code and computer use applications, operate under disclosed prompt injection risk that the document states is partially but not fully mitigated by applied safeguards. The bug bounty program provides a mechanism for security researchers to report discovered vulnerabilities.

Is ConductAtlas affiliated with Anthropic?

No. ConductAtlas is an independent monitoring service. We are not affiliated with, endorsed by, or sponsored by Anthropic.