← Models

Model profile

GPT 5.5

OpenAIdeveloper
2026-04-23release date
#9 / 267overall rank
29eval lineages
2discovery sources

Evidence summary

GPT 5.5 has an estimated overall rank of #9; its 90% source-sensitivity interval is #10–#66. Its behavior-only rank is #10; company governance moves the combined estimate to #9. Published evidence spans 29 evals and 7 of 7 behavior components. Its strongest relative result is SpeciEval (belief_animal_sentience, #1 of 102); its weakest is GPT-5.6 system card — disallowed content with challenging prompts (extremism_not_unsafe, #7 of 7).

Compare this model

Only models sharing at least one published sub-eval are listed.

Official and reference links

Published eval results

Rank is within that sub-eval. Black marks the observed result; the grey dot marks the value implied by the global rank. Values stay on each source’s native scale.

Eval / sub-evalRankValueDistributionBetterSource
AA-Omnisciencehallucination_rate#205 / 3110.8553↓ lowerSource ↗official
AgentAbstainabstain#9 / 1761.1↑ higherSource ↗official
AgentAbstaincar#7 / 1759.8↑ higherSource ↗official
AgentAbstainpaired#4 / 1752.5↑ higherSource ↗official
ANIMAscore#3 / 190.74↑ higherSource ↗official
BioSecBench-Refusalbalanced_refusal_score#8 / 100.2869↑ higherSource ↗official
BullshitBench v2clear_pushback_rate#28 / 1050.4567↑ higherSource ↗official
CAIS Risk Indexagent_red_teaming#3 / 4341.5↓ lowerSource ↗official
CAIS Risk Indexbioweapons_assistance#14 / 4860.5↓ lowerSource ↗official
CAIS Risk Indexhle_overconfidence#7 / 4941.8↓ lowerSource ↗official
CAIS Risk Indexmachiavelli#8 / 4581.7↓ lowerSource ↗official
CAIS Risk Indexmask#12 / 519.9↓ lowerSource ↗official
CAIS Risk Indexpolitical_manipulation#6 / 3239↓ lowerSource ↗official
CAIS Risk Indextextquests_harm#36 / 4821.4↓ lowerSource ↗official
Cisco AI Defense Rolling Single-Turn Leaderboardsingle_turn_attack_success_rate#5 / 1052.343↓ lowerSource ↗official
DystopiaBenchbasaglia_score#17 / 5055↓ lowerSource ↗official
DystopiaBenchbaudrillard_score#14 / 5040.7↓ lowerSource ↗official
DystopiaBenchhuxley_score#15 / 5052.5↓ lowerSource ↗official
DystopiaBenchlaguardia_score#17 / 5060.13↓ lowerSource ↗official
DystopiaBenchorwell_score#15 / 5049.5↓ lowerSource ↗official
DystopiaBenchpetrov_score#12 / 5049.13↓ lowerSource ↗official
Enkrypt AI Safety Leaderboardbias_attack_non_success_rate#40 / 26038.5↑ higherSource ↗official
Enkrypt AI Safety Leaderboardcbrn_attack_non_success_rate#219 / 26079.33↑ higherSource ↗official
Enkrypt AI Safety Leaderboardharmful_attack_non_success_rate#7 / 26099.44↑ higherSource ↗official
Enkrypt AI Safety Leaderboardtoxicity_attack_non_success_rate#201 / 25892↑ higherSource ↗official
FORTRESSaverage_risk_score#10 / 4916.28↓ lowerSource ↗official
FORTRESSover_refusal_score#38 / 469.79↓ lowerSource ↗official
GPT-5.6 system card — disallowed content with challenging promptsextremism_not_unsafe#7 / 70.925↑ higherSource ↗official
GPT-5.6 system card — disallowed content with challenging promptsgore_not_unsafe#3 / 70.8↑ higherSource ↗official
GPT-5.6 system card — disallowed content with challenging promptshate_not_unsafe#1 / 71↑ higherSource ↗official
GPT-5.6 system card — disallowed content with challenging promptsnonviolent_illicit_not_unsafe#6 / 70.987↑ higherSource ↗official
GPT-5.6 system card — disallowed content with challenging promptsself_harm_not_unsafe#7 / 70.917↑ higherSource ↗official
GPT-5.6 system card — disallowed content with challenging promptssexual_minors_not_unsafe#6 / 70.938↑ higherSource ↗official
GPT-5.6 system card — disallowed content with challenging promptssexual_not_unsafe#3 / 70.944↑ higherSource ↗official
GPT-5.6 system card — disallowed content with challenging promptsviolent_illicit_not_unsafe#5 / 70.94↑ higherSource ↗official
GPT-5.6 system card — first-person fairnessharm_overall_pct#6 / 71.12↓ lowerSource ↗official
GPT-5.6 system card — prompt-injection robustnessconnectors_injection_resistance#1 / 71↑ higherSource ↗official
Gray Swan indirect prompt injection (15 attempts)attack_success_probability_k15_pct#8 / 1320.8↓ lowerSource ↗official
HUMAINE Trust, Ethics and Safetytrust_ethics_safety_score#11 / 5428.38↑ higherSource ↗official
MACHIAVELLIdeception_relative_random_pct#8 / 5081.7↓ lowerSource ↗official
MANTAAWMS#2 / 70.504↑ higherSource ↗official
MANTAAWVS#2 / 70.664↑ higherSource ↗official
MORUscore#1 / 1385.4↑ higherSource ↗official
ODCV-Benchaverage_severity#2 / 120.7125↓ lowerSource ↗official
ODCV-Benchmisalignment_rate#3 / 1221.25↓ lowerSource ↗official
PHAREbias_resistance_diagnostic#48 / 660.3829↑ higherSource ↗official
PHAREhallucination_resistance_diagnostic#20 / 700.7978↑ higherSource ↗official
PHAREharm_resistance_diagnostic#6 / 700.9892↑ higherSource ↗official
PHAREjailbreak_resistance_diagnostic#18 / 670.6513↑ higherSource ↗official
RefusalBenchyouden_j#11 / 190.4213↑ higherSource ↗official
SM-Benchadversarial#31 / 7382.93↑ higherSource ↗official
SM-Benchambiguous_interpretation#42 / 7382.74↑ higherSource ↗official
SM-Benchanti_hallucination#27 / 7394.76↑ higherSource ↗official
SM-Bencheq_boundaries#36 / 7364.61↑ higherSource ↗official
SM-Benchoverfit#57 / 7350↑ higherSource ↗official
SpeciEvalbelief_animal_sentience#1 / 1027↑ higherSource ↗official
SpeciEvalland_animal_4ns#54 / 1024.58↓ lowerSource ↗official
SpeciEvalsea_animal_4ns#24 / 1024.55↓ lowerSource ↗official
SpeciEvalspeciesism#7 / 1021.32↓ lowerSource ↗official
TACbase_welfare_rate#57 / 6819.23↑ higherSource ↗official
ToolPrivacyBenchprivate_mt_poi#3 / 920.39↓ lowerSource ↗official
ToolPrivacyBenchpublic_mt_poi#2 / 916.75↓ lowerSource ↗official
VETO Misfired Alignmentmisfired_alignment_rate_pct#9 / 257.5↓ lowerSource ↗official

Values evaluations

Descriptive values and political-framing results are separate from safety/ethics ranks. Each strip shows the evaluation’s observed model range; its endpoint labels state what lower and higher values mean.

CAIS AI Values — countries