Gemini 1.5
Google DeepMind — Gemini
rich detail 13 method fields published
Evaluation
MTOB benchmark (Tanzer et al., 2023): learn to translate English-Kalamang (ISO kgv, <200 speakers) from a ~500-page grammar and ~2000-entry wordlist in context; human evaluation with 0-6 quality scale; compared to GPT-4 Turbo and Claude 3 (half-grammar) and human learner
How it was run
Open a release to see the setup, scoring, and other details that the lab published. Only fields the lab actually disclosed are shown.
rich detail 13 method fields published