Gemma 1
Google DeepMind — Gemma
rich detail 18 method fields published
Evaluation
representational harms benchmarked against academic datasets such as WinoBias and BBQ; safety benchmark results additionally reported on BBQ, BOLD, Winogender, Winobias, RealToxicity and TruthfulQA
How it was run
Open a release to see the setup, scoring, and other details that the lab published. Only fields the lab actually disclosed are shown.
rich detail 18 method fields published
rich detail 18 method fields published