o1
OpenAI
rich detail 10 method fields published
Evaluation
Refusals for multimodal inputs on standard set for disallowed text+image content and overrefusals (categories: sexual/exploitative, self-harm/intent, self-harm/instructions)
How it was run
Open a release to see the setup, scoring, and other details that the lab published. Only fields the lab actually disclosed are shown.
rich detail 10 method fields published
rich detail 10 method fields published