Skip to content

Evaluation

Cyber attack uplift study and attack automation study

Cyber attack uplift study (whether LLMs enhance human hacking capability in skill level and speed) and attack automation study (LLMs as autonomous agents in ransomware attacks without human intervention)

Cybersecurity Meta

How it was run

Open a release to see the setup, scoring, and other details that the lab published. Only fields the lab actually disclosed are shown.

Llama 3.1
Meta

partial detail 6 method fields published

See how this test was run