Llama 3.1
Meta
partial detail 8 method fields published
Evaluation
Uplift testing to assess whether Llama 3.1 models could meaningfully increase the capabilities of malicious actors to plan or carry out attacks using chemical and biological weapons
How it was run
Open a release to see the setup, scoring, and other details that the lab published. Only fields the lab actually disclosed are shown.
partial detail 8 method fields published