Skip to content

Evaluation

Standard benchmark suite (zero-shot and few-shot, 20 benchmarks)

Zero-shot and few-shot performance of LLaMA models on 20 standard benchmarks

General benchmark suites Meta

How it was run

Open a release to see the setup, scoring, and other details that the lab published. Only fields the lab actually disclosed are shown.

Llama 1
Meta

partial detail 16 method fields published

See how this test was run