Provision record
Meta · Llama 4 Model Card · View original document ↗

Red Teaming and Critical Risk Evaluation

Low severity High confidence Explicit document language Unique · 0 of 352 platforms
Stay ahead of the changes
Track Meta and get the diff the day its terms change.
Share 𝕏 Share in Share 🔒 PDF
Document Record

What it is

The card discloses that Meta conducted recurring red teaming exercises and targeted evaluations across three critical risk categories, CBRNE, child safety, and cyber attack enablement, and describes the methodology and scope of those evaluations.

ⓘ

This analysis describes what Meta's agreement states, permits, or reserves. It does not constitute a legal determination about enforceability. Regulatory applicability and practical outcomes may vary by jurisdiction, enforcement context, and individual circumstances. Read our methodology

ConductAtlas Analysis

Why it matters (compliance & governance perspective)

This provision documents the scope and methodology of Meta's pre-release safety evaluation for Llama 4, providing institutional deployers with a basis for assessing the categories of risk that were formally evaluated prior to release and those that were not.

Consumer impact (what this means for users)

The document states that safety evaluations covered CBRNE-related prompts, child safety outputs, and cyberattack enablement capabilities, and that these evaluations informed safety fine-tuning. The card also notes that testing could not cover all scenarios and that developers should conduct additional testing for their specific applications.

Cross-platform context

See how other platforms handle Red Teaming and Critical Risk Evaluation and similar clauses.

Compare across platforms →
▸ View Original Clause Language DOCUMENT RECORD
"
We conduct recurring red teaming exercises with the goal of discovering risks via adversarial prompting and we use the learnings to improve our benchmarks and safety tuning datasets. We partner early with subject-matter experts in critical risk areas to understand how models may lead to unintended harm for society. We spend additional focus on the following critical risk areas: CBRNE (Chemical, Biological, Radiological, Nuclear, and Explosive materials) helpfulness... Child Safety... Cyber attack enablement.

Excerpt from Meta's Llama 4 Model Card

ConductAtlas Analysis

Institutional analysis (regulatory & governance intelligence)

(1) REGULATORY LANDSCAPE: The disclosure of CBRNE and child safety evaluations engages potential obligations under the EU AI Act for transparency about safety testing methodologies for general-purpose AI models.

Insight

Unlock the full institutional analysis

Enforcement risk, jurisdiction flags, contract triggers, and due diligence action items.

Applicable agencies

  • Federal Trade Commission (ftc)
    Oversees unfair or deceptive business practices and can investigate companies that mislead consumers about data collection, sharing, or use.
    Who can file: Anyone affected by the company's practices (US or international)
    What you need: Your account details, a timeline of relevant events, and a description of the specific issue
    What to expect: Complaints inform FTC enforcement priorities and investigations but do not result in individual resolution or compensation
    File a complaint →

Provision details

Document information
Document
Llama 4 Model Card
Entity
Meta
Document last updated
July 6, 2026
Tracking information
First tracked
July 6, 2026
Last verified
July 6, 2026
Record ID
CA-P-013386
Document ID
CA-D-00922
Evidence Provenance
Source URL
Wayback Machine
Content hash (SHA-256)
a1739083c8a496ca0c121d90b98659a6aa7e36fcc0092581db35ef99736d104f
Analysis generated
July 6, 2026 22:02 UTC
Methodology
Evidence
✓ Snapshot stored   ✓ Hash verified
Citation Record
Entity: Meta
Document: Llama 4 Model Card
Record ID: CA-P-013386
Captured: 2026-07-06 22:02:11 UTC
SHA-256: a1739083c8a496ca…
URL: https://conductatlas.com/platform/meta/llama-4-model-card/provision/CA-P-013386/red-teaming-and-critical-risk-evaluation/
Accessed: Oct. 2, 2026
Permanent archival reference. Stable identifier suitable for legal filings, compliance documentation, and research citation.
Classification
Severity
Low
Categories

Other risks in this policy

Get the research letter

Companies change their terms quietly. We read every version and catch what actually changed. One email a week on the changes that matter and what they mean.

Frequently Asked Questions

What does Meta's Red Teaming and Critical Risk Evaluation clause do?

This provision documents the scope and methodology of Meta's pre-release safety evaluation for Llama 4, providing institutional deployers with a basis for assessing the categories of risk that were formally evaluated prior to release and those that were not.

How does this clause affect you?

The document states that safety evaluations covered CBRNE-related prompts, child safety outputs, and cyberattack enablement capabilities, and that these evaluations informed safety fine-tuning. The card also notes that testing could not cover all scenarios and that developers should conduct additional testing for their specific applications.

Is ConductAtlas affiliated with Meta?

No. ConductAtlas is an independent monitoring service. We are not affiliated with, endorsed by, or sponsored by Meta.