Llama 3.1
Meta
partial detail 9 method fields published
Evaluation
Child safety risk assessments by a team of experts across multiple attack vectors (including Llama 3 languages), with content specialists; findings inform fine-tuning mitigations and expand evaluation benchmark coverage
How it was run
Open a release to see the setup, scoring, and other details that the lab published. Only fields the lab actually disclosed are shown.
partial detail 9 method fields published
partial detail 9 method fields published
partial detail 9 method fields published
partial detail 9 method fields published