o1-preview
OpenAI
rich detail 11 method fields published
Evaluation
Benign prompts from XSTest testing over-refusal edge cases (categories: Definitions, Figurative Language, Historical Events, Homonyms, Discr: Nonsense group, Discr: Nonsense context, Privacy: fictional, Privacy: public, Safe Contexts, Safe Targets)
How it was run
Open a release to see the setup, scoring, and other details that the lab published. Only fields the lab actually disclosed are shown.
rich detail 11 method fields published
rich detail 11 method fields published
rich detail 11 method fields published
rich detail 11 method fields published