← Back to Explorer
Evaluation 366 of 723
Evaluation
Internal red-teaming of content policies
internal red-teaming testing of relevant content policies, conducted by a number of different teams each with different goals and human evaluation metrics
Red teaming & adversarial testing
Google DeepMind — Gemma
How it was run
Open a release to see the setup, scoring, and other details that the lab published. Only fields the lab actually disclosed are shown.
Gemma 1
Feb 2024
Google DeepMind — Gemma
partial detail
17 method fields published
See how this test was run
Gemma model card | Google AI for Developers · Feb 2024
Gemma model card | Google AI for Developers · Feb 2024
Gemma model card | Google AI for Developers · Feb 2024
CodeGemma
Apr 2024
Google DeepMind — Gemma
partial detail
17 method fields published
See how this test was run
CodeGemma model card | Google AI for Developers · Apr 2024
Gemma 2
Jun 2024
Google DeepMind — Gemma
partial detail
17 method fields published
See how this test was run
Gemma 2 model card | Google AI for Developers · Jun 2024
Gemma 2 model card | Google AI for Developers · Jun 2024
PaliGemma 1
Jul 2024
Google DeepMind — Gemma
partial detail
17 method fields published
See how this test was run
PaliGemma 1 model card | Google AI for Developers · Jul 2024
PaliGemma 1 model card | Google AI for Developers · Jul 2024
Gemma 3
Mar 2025
Google DeepMind — Gemma
partial detail
17 method fields published
See how this test was run
Gemma 3 model card | Google AI for Developers · Mar 2025
Gemma 3 model card | Google AI for Developers · Mar 2025
Gemma 3n
May 2025
Google DeepMind — Gemma
partial detail
18 method fields published
See how this test was run
Gemma 3n model card | Google AI for Developers · May 2025
Gemma 3n model card | Google AI for Developers · May 2025