Skip to content

Evaluation

Safe completions safety/helpfulness evaluation

Safety and helpfulness across prompt intent types with safe-completions training

Core capabilities OpenAI

How it was run

Open a release to see the setup, scoring, and other details that the lab published. Only fields the lab actually disclosed are shown.

GPT-5
OpenAI

rich detail 7 method fields published

See how this test was run