Skip to content

Evaluation

Long-document QA (Les Misérables)

QA on the full 1,462-page Les Misérables (710K tokens) in context, with side-by-side comparisons vs Gemini 1.0 Pro using retrieval-augmented generation (TF-IDF indexing, 4k token retrieval)

Core capabilities Google DeepMind — Gemini

How it was run

Open a release to see the setup, scoring, and other details that the lab published. Only fields the lab actually disclosed are shown.

Gemini 1.5
Google DeepMind — Gemini

rich detail 14 method fields published

See how this test was run