Chaptered summary is still being generated for this document. Showing a heuristic brief in the meantime.
Summary
17,527-word document condensed to 121 words. OpenAI · Aug 20, 2026
TL;DR
“The o1 model series is trained with large-scale reinforcement learning to reason using chain of thought. These advanced reasoning capabilities provide new avenues for improving the safety and robustness of our models.”
Top benchmarks
| Benchmark | Variant | Score |
|---|---|---|
| SWE-bench | pre-mitigation, verified, pass_at_1 | 41.3% |
| SWE-bench | post-mitigation, verified, pass_at_1 | 41.3% |
| SWE-bench | post-mitigation, verified, pass_at_1 | 40.9% |
Showing top 3 of 26. See full list below.
Capability claim
- “We trained the summarizer model away from producing disallowed content in these summaries.”
Safety findings
- “not release in products) are denoted as “pre-mitigation,” specifically o1 (pre-mitigation).”
Deployment scope
- “available to us, we believe the post-mitigation o1 model cannot meaningfully assist in the development of radiological or nuclear weapons, but note again that this assessment is limited by what we can test.”
Every italicized passage is a verbatim substring of the source document (checked deterministically after extraction). Field selection is heuristic — some quotes may lack surrounding context and some claims may be absent if no matching pattern appeared. For citation, open the source: original model card · source SHA 4d4a0338953d · version dated Aug 20, 2026.
Extracted Evaluations(26 results)
Sort by:⚠ 1 conflicting report0/26 rows fully reproducible (0%)
| Benchmark | Category | State | Score | Setup | Source |
|---|---|---|---|---|---|
/ verified | coding | scored | 41.3% pass at 1 | pre-mitigationmissing: shot countmissing: languagemissing: training state | self-reported |
/ verified | coding | scored | 41.3% pass at 1 | post-mitigationmissing: shot countmissing: languagemissing: training state | self-reported |
/ verified⚠ 1 other disagree | coding | scored | 40.9% pass at 1 | post-mitigationmissing: shot countmissing: languagemissing: training state | self-reported |
| coding | cited | — | missing: shot countmissing: methodmissing: languagemissing: training state | self-reported | |
Agentic Tasks | other | mentioned | — | missing: shot countmissing: methodmissing: languagemissing: training state | self-reported |
| other | mentioned | — | missing: shot countmissing: methodmissing: languagemissing: training state | self-reported | |
| other | mentioned | — | missing: shot countmissing: methodmissing: languagemissing: training state | self-reported | |
| other | mentioned | — | missing: shot countmissing: methodmissing: languagemissing: training state | self-reported | |
| other | mentioned | — | missing: shot countmissing: methodmissing: languagemissing: training state | self-reported | |
| other | mentioned | — | missing: shot countmissing: methodmissing: languagemissing: training state | self-reported | |
| other | mentioned | — | missing: shot countmissing: methodmissing: languagemissing: training state | self-reported | |
| other | mentioned | — | missing: shot countmissing: methodmissing: languagemissing: training state | self-reported | |
Wildchat/ toxic | other | mentioned | — | missing: shot countmissing: methodmissing: languagemissing: training state | self-reported |
Wildchat/ toxic | other | mentioned | — | missing: shot countmissing: methodmissing: languagemissing: training state | self-reported |
Wildchat/ toxic | other | mentioned | — | missing: shot countmissing: methodmissing: languagemissing: training state | self-reported |
Wildchat/ toxic | other | mentioned | — | missing: shot countmissing: methodmissing: languagemissing: training state | self-reported |
OpenAI Research Engineer Interview/ coding | other | mentioned | — pass at 1 | missing: shot countmissing: methodmissing: languagemissing: training state | self-reported |
| other | mentioned | — | missing: shot countmissing: methodmissing: languagemissing: training state | self-reported | |
| other | mentioned | — | missing: shot countmissing: methodmissing: languagemissing: training state | self-reported | |
OpenAI Research Engineer Interview/ multiple_choice | other | mentioned | — | missing: shot countmissing: methodmissing: languagemissing: training state | self-reported |
OpenAI Research Engineer Interview/ coding | other | mentioned | — pass at 128 | missing: shot countmissing: methodmissing: languagemissing: training state | self-reported |
OpenAI Research Engineer Interview/ multiple_choice | other | mentioned | — cons at 32 | majority-votingmissing: shot countmissing: languagemissing: training state | self-reported |
| safety | mentioned | — | missing: shot countmissing: methodmissing: languagemissing: training state | self-reported | |
| safety | mentioned | — | missing: shot countmissing: methodmissing: languagemissing: training state | self-reported | |
| safety | mentioned | — | missing: shot countmissing: methodmissing: languagemissing: training state | self-reported | |
| safety | mentioned | — | missing: shot countmissing: methodmissing: languagemissing: training state | self-reported |