Model Cards / OpenAI

GPT-5.6 System Card

model card21,019 words·91 min read·Aug 3, 2026·Source
Chaptered summary is still being generated for this document. Showing a heuristic brief in the meantime.
Summary
21,019-word document condensed to 150 words. OpenAI · Aug 3, 2026
TL;DR

GPT-5.6 is a new family of three models: Sol, our new flagship model; Terra, a capable lower-cost option; and Luna, our fastest and most cost-efficient model. The safeguards we have built for this launch—our most robust yet— are built to deliver these models safely and at scale, around the world.

Top benchmarks
BenchmarkVariantScore
Tacit Knowledge and Troubleshootingmcq, accuracy_with_refusal_adjustment84.1%
Tacit Knowledge and Troubleshootingmcq, accuracy_with_refusal_adjustment83.8%
Tacit Knowledge and Troubleshootingmcq, accuracy65.0%
Multimodal Troubleshooting Virologyaccuracy55.5%
TroubleshootingBenchaccuracy48.0%
ProtocolQAopen_ended, accuracy43.5%
DNA Sequence Design for Transcription Factor Bindingpass_at_116.5%
DNA Sequence Design for Transcription Factor Bindingpass_at_113.8%

Showing top 8 of 50. See full list below.

Capability claim
  • we trained the models to maintain a strong standard of overwrite avoidance while improving autonomy without relying on extra cautious prompting.
Mitigations
  • We have deployed an expanded set of safeguards to restrict the ability of malicious actors to benefit from increased capabilities in cybersecurity performance.
  • we have deployed Preparedness Safeguards.
Deployment scope
  • available to the public, we can continue to reserve the most sensitive cybersecurity and biological capabilities for trusted defenders.

Every italicized passage is a verbatim substring of the source document (checked deterministically after extraction). Field selection is heuristic — some quotes may lack surrounding context and some claims may be absent if no matching pattern appeared. For citation, open the source: original model card · source SHA a273db38c41c · version dated Aug 3, 2026.

Extracted Evaluations(50 results)

Sort by:0/50 rows fully reproducible (0%)
BenchmarkCategoryStateScoreSetupSource
codingcited
missing: shot countmissing: methodmissing: languagemissing: training state
self-reported
/ verified
codingcited
missing: shot countmissing: methodmissing: languagemissing: training state
self-reported
/ pro
knowledgecited
missing: shot countmissing: methodmissing: languagemissing: training state
self-reported
Tacit Knowledge and Troubleshooting/ mcq
otherscored
84.1
accuracy with refusal adjustment
missing: shot countmissing: methodmissing: languagemissing: training state
self-reported
Tacit Knowledge and Troubleshooting/ mcq
otherscored
83.8
accuracy with refusal adjustment
missing: shot countmissing: methodmissing: languagemissing: training state
self-reported
Tacit Knowledge and Troubleshooting/ mcq
otherscored
65.0
accuracy
missing: shot countmissing: methodmissing: languagemissing: training state
self-reported
otherscored
55.5
accuracy
missing: shot countmissing: methodmissing: languagemissing: training state
self-reported
otherscored
48.0
accuracy
missing: shot countmissing: methodmissing: languagemissing: training state
self-reported
/ open_ended
otherscored
43.5
accuracy
missing: shot countmissing: methodmissing: languagemissing: training state
self-reported
otherscored
16.5
pass at 1
missing: shot countmissing: methodmissing: languagemissing: training state
self-reported
otherscored
13.8
pass at 1
missing: shot countmissing: methodmissing: languagemissing: training state
self-reported
otherscored
13.7
pass at 1
missing: shot countmissing: methodmissing: languagemissing: training state
self-reported
otherscored
12.8
pass at 1
missing: shot countmissing: methodmissing: languagemissing: training state
self-reported
otherscored
7.6
pass at 4
missing: shot countmissing: methodmissing: languagemissing: training state
self-reported
otherscored
3.5
pass at 4
missing: shot countmissing: methodmissing: languagemissing: training state
self-reported
/ 5k_tokens
otherscored
1.3
control rate
missing: shot countmissing: methodmissing: languagemissing: training state
self-reported
/ 5k_tokens
otherscored
0.7
control rate
missing: shot countmissing: methodmissing: languagemissing: training state
self-reported
AAV_Capsid_Packaging_Prediction
otherscored
0.5
spearman correlation
missing: shot countmissing: methodmissing: languagemissing: training state
self-reported
AAV_Capsid_Packaging_Prediction
otherscored
0.5
spearman correlation
missing: shot countmissing: methodmissing: languagemissing: training state
self-reported
otherscored
0.4
pass at 4
missing: shot countmissing: methodmissing: languagemissing: training state
self-reported
/ 5k_tokens
otherscored
0.4
control rate
missing: shot countmissing: methodmissing: languagemissing: training state
self-reported
AAV_Capsid_Packaging_Prediction
otherscored
0.3
spearman correlation
missing: shot countmissing: methodmissing: languagemissing: training state
self-reported
otherscored
0.0
pass at 4
missing: shot countmissing: methodmissing: languagemissing: training state
self-reported
othercited
missing: shot countmissing: methodmissing: languagemissing: training state
self-reported
Scruples
othermentioned
missing: shot countmissing: methodmissing: languagemissing: training state
self-reported
Scruples/ suggest_right
othermentioned
g mean 2
missing: shot countmissing: methodmissing: languagemissing: training state
self-reported
Scruples/ first_person
othermentioned
missing: shot countmissing: methodmissing: languagemissing: training state
self-reported
Flaky Tools
othermentioned
g mean 2
missing: shot countmissing: methodmissing: languagemissing: training state
self-reported
Impossible Coding Tasks
othermentioned
g mean 2
missing: shot countmissing: methodmissing: languagemissing: training state
self-reported
Honesty
othermentioned
missing: shot countmissing: methodmissing: languagemissing: training state
self-reported
Model Spec
othermentioned
missing: shot countmissing: methodmissing: languagemissing: training state
self-reported
Instruction Hierarchy
othermentioned
missing: shot countmissing: methodmissing: languagemissing: training state
self-reported
Deployment Simulation
othermentioned
missing: shot countmissing: methodmissing: languagemissing: training state
self-reported
First-Person Fairness
othermentioned
harm overall
missing: shot countmissing: methodmissing: languagemissing: training state
self-reported
othermentioned
accuracy
missing: shot countmissing: methodmissing: languagemissing: training state
self-reported
othermentioned
accuracy
missing: shot countmissing: methodmissing: languagemissing: training state
self-reported
/ open_ended
othermentioned
accuracy
missing: shot countmissing: methodmissing: languagemissing: training state
self-reported
/ open_ended
othermentioned
accuracy
missing: shot countmissing: methodmissing: languagemissing: training state
self-reported
Tacit Knowledge and Troubleshooting/ mcq
othermentioned
accuracy with refusal adjustment
missing: shot countmissing: methodmissing: languagemissing: training state
self-reported
Tacit Knowledge and Troubleshooting/ mcq
othermentioned
accuracy with refusal adjustment
missing: shot countmissing: methodmissing: languagemissing: training state
self-reported
Tacit Knowledge and Troubleshooting/ mcq
othermentioned
accuracy with refusal adjustment
missing: shot countmissing: methodmissing: languagemissing: training state
self-reported
othermentioned
accuracy
missing: shot countmissing: methodmissing: languagemissing: training state
self-reported
othermentioned
accuracy
missing: shot countmissing: methodmissing: languagemissing: training state
self-reported
othermentioned
accuracy
missing: shot countmissing: methodmissing: languagemissing: training state
self-reported
Agentic Misalignment
othermentioned
g mean 2
missing: shot countmissing: methodmissing: languagemissing: training state
self-reported
Agentic Misalignment/ destructive_actions
othermentioned
g mean 2
missing: shot countmissing: methodmissing: languagemissing: training state
self-reported
Agentic Misalignment/ background_work
othermentioned
g mean 2
missing: shot countmissing: methodmissing: languagemissing: training state
self-reported
Health Queries/ patient_opinion
othermentioned
tpr
missing: shot countmissing: methodmissing: languagemissing: training state
self-reported
Impossible Tasks
othermentioned
missing: shot countmissing: methodmissing: languagemissing: training state
self-reported
reasoningcited
missing: shot countmissing: methodmissing: languagemissing: training state
self-reported