Skip to content

Evaluation

CBRNE uplift testing

Uplift testing to assess whether Llama 3.1 models could meaningfully increase the capabilities of malicious actors to plan or carry out attacks using chemical and biological weapons

Biological & chemical risks Meta

How it was run

Open a release to see the setup, scoring, and other details that the lab published. Only fields the lab actually disclosed are shown.

Llama 3.1
Meta

partial detail 8 method fields published

See how this test was run