Gemma 1
Google DeepMind — Gemma
rich detail 24 method fields published
Evaluation
win rate of Gemma IT models versus Mistral v0.2 7B Instruct on a held-out collection of ~1000 prompts (creative writing, coding, instruction following) and ~400 prompts testing basic safety protocols
How it was run
Open a release to see the setup, scoring, and other details that the lab published. Only fields the lab actually disclosed are shown.
rich detail 24 method fields published
partial detail 20 method fields published