Llama 2
Meta
rich detail 18 method fields published
Evaluation
Human ratings of major Llama 2-Chat model versions on helpfulness and safety vs open-source (Falcon, MPT, Vicuna) and closed-source (ChatGPT, PaLM) models on over 4,000 single- and multi-turn prompts
How it was run
Open a release to see the setup, scoring, and other details that the lab published. Only fields the lab actually disclosed are shown.
rich detail 18 method fields published