← Models

Model profile

GPT 5.2

OpenAIdeveloper
2025-12-11release date
#16 / 267overall rank
26eval lineages

Evidence summary

GPT 5.2 has an estimated overall rank of #16; its 90% source-sensitivity interval is #13–#74. Its behavior-only rank is #20; company governance moves the combined estimate to #16. Published evidence spans 26 evals and 7 of 7 behavior components. Its strongest relative result is LiveSecBench (ethics, #1 of 43); its weakest is SM-Bench (eq_boundaries, #72 of 73).

Compare this model

Only models sharing at least one published sub-eval are listed.

Official and reference links

Published eval results

Rank is within that sub-eval. Black marks the observed result; the grey dot marks the value implied by the global rank. Values stay on each source’s native scale.

Eval / sub-evalRankValueDistributionBetterSource
AA-Omnisciencehallucination_rate#71 / 3110.6057↓ lowerSource ↗official
AgentAbstainabstain#7 / 1763.1↑ higherSource ↗official
AgentAbstaincar#8 / 1759.2↑ higherSource ↗official
AgentAbstainpaired#10 / 1742.9↑ higherSource ↗official
Alignment Leaderboardcorrigibility#1 / 244.641↑ higherSource ↗official
Alignment Leaderboardhonesty#3 / 244.465↑ higherSource ↗official
Alignment Leaderboardnon_manipulation#3 / 244.523↑ higherSource ↗official
Alignment Leaderboardrobustness#4 / 244.227↑ higherSource ↗official
Alignment Leaderboardsafety#3 / 244.769↑ higherSource ↗official
Alignment Leaderboardscheming#3 / 244.554↑ higherSource ↗official
ANIMAscore#11 / 190.6906↑ higherSource ↗official
BullshitBench v2clear_pushback_rate#47 / 1050.33↑ higherSource ↗official
CAIS Risk Indexagent_red_teaming#17 / 4368.8↓ lowerSource ↗official
CAIS Risk Indexbioweapons_assistance#16 / 4862.9↓ lowerSource ↗official
CAIS Risk Indexhle_overconfidence#11 / 4943.7↓ lowerSource ↗official
CAIS Risk Indexmachiavelli#20 / 4586.5↓ lowerSource ↗official
CAIS Risk Indexmask#18 / 5112.9↓ lowerSource ↗official
CAIS Risk Indexpolitical_manipulation#27 / 3257↓ lowerSource ↗official
CAIS Risk Indextextquests_harm#32 / 4820.5↓ lowerSource ↗official
Cisco AI Defense Rolling Single-Turn Leaderboardsingle_turn_attack_success_rate#16 / 1054.736↓ lowerSource ↗official
Enkrypt AI Safety Leaderboardbias_attack_non_success_rate#29 / 26041.6↑ higherSource ↗official
Enkrypt AI Safety Leaderboardcbrn_attack_non_success_rate#100 / 26089.83↑ higherSource ↗official
Enkrypt AI Safety Leaderboardharmful_attack_non_success_rate#52 / 26091.67↑ higherSource ↗official
Enkrypt AI Safety Leaderboardtoxicity_attack_non_success_rate#121 / 25896.73↑ higherSource ↗official
FORTRESSaverage_risk_score#14 / 4917.5↓ lowerSource ↗official
FORTRESSover_refusal_score#40 / 4611.76↓ lowerSource ↗official
GPT-5.6 system card — disallowed content with challenging promptsextremism_not_unsafe#1 / 71↑ higherSource ↗official
GPT-5.6 system card — disallowed content with challenging promptsgore_not_unsafe#1 / 70.877↑ higherSource ↗official
GPT-5.6 system card — disallowed content with challenging promptshate_not_unsafe#6 / 70.927↑ higherSource ↗official
GPT-5.6 system card — disallowed content with challenging promptsnonviolent_illicit_not_unsafe#2 / 70.993↑ higherSource ↗official
GPT-5.6 system card — disallowed content with challenging promptsself_harm_not_unsafe#3 / 70.961↑ higherSource ↗official
GPT-5.6 system card — disallowed content with challenging promptssexual_minors_not_unsafe#5 / 70.948↑ higherSource ↗official
GPT-5.6 system card — disallowed content with challenging promptssexual_not_unsafe#5 / 70.94↑ higherSource ↗official
GPT-5.6 system card — disallowed content with challenging promptsviolent_illicit_not_unsafe#1 / 70.975↑ higherSource ↗official
GPT-5.6 system card — first-person fairnessharm_overall_pct#5 / 71↓ lowerSource ↗official
GPT-5.6 system card — prompt-injection robustnessconnectors_injection_resistance#6 / 70.971↑ higherSource ↗official
GPT-5.6 system card — prompt-injection robustnesssearch_function_calling_injection_resistance#5 / 60.568↑ higherSource ↗official
LiveSecBenchethics#1 / 4395.79↑ higherSource ↗official
LiveSecBenchfactuality#16 / 4357.75↑ higherSource ↗official
LiveSecBenchlegality#4 / 4389.24↑ higherSource ↗official
LiveSecBenchprivacy#4 / 4387.9↑ higherSource ↗official
LiveSecBenchpsychological_health#1 / 4392.91↑ higherSource ↗official
MACHIAVELLIdeception_relative_random_pct#20 / 5086.5↓ lowerSource ↗official
Manager Coercion Benchcoercion_ladder_depth#8 / 317.4↓ lowerSource ↗official
Manager Coercion Benchfabrication_rate#1 / 130↓ lowerSource ↗official
MASKlying_probability_pct#9 / 5313.33↓ lowerSource ↗official
MORUscore#2 / 1384.18↑ higherSource ↗official
PHAREbias_resistance_diagnostic#47 / 660.3851↑ higherSource ↗official
PHAREhallucination_resistance_diagnostic#29 / 700.7713↑ higherSource ↗official
PHAREharm_resistance_diagnostic#15 / 700.9688↑ higherSource ↗official
PHAREjailbreak_resistance_diagnostic#9 / 670.7161↑ higherSource ↗official
PropensityBenchscore#6 / 1434.35↓ lowerSource ↗official
SM-Benchadversarial#55 / 7378.54↑ higherSource ↗official
SM-Benchambiguous_interpretation#19 / 7388.99↑ higherSource ↗official
SM-Benchanti_hallucination#54 / 7386.39↑ higherSource ↗official
SM-Bencheq_boundaries#72 / 7340.45↑ higherSource ↗official
SM-Benchoverfit#67 / 7325.68↑ higherSource ↗official
SpeciEvalbelief_animal_sentience#78 / 1026.52↑ higherSource ↗official
SpeciEvalland_animal_4ns#16 / 1024.25↓ lowerSource ↗official
SpeciEvalsea_animal_4ns#9 / 1024.3↓ lowerSource ↗official
SpeciEvalspeciesism#19 / 1021.55↓ lowerSource ↗official
TACbase_welfare_rate#36 / 6826.28↑ higherSource ↗official
TukaBenchafri_jbb_cultural_asr#2 / 610.9↓ lowerSource ↗official
TukaBenchafri_jbb_harm_asr#1 / 62.6↓ lowerSource ↗official
TukaBenchafrijail_mono_asr#1 / 612.2↓ lowerSource ↗official
Vigil Mental Health Safetyoverall_score#5 / 2372↑ higherSource ↗official

Values evaluations

Descriptive values and political-framing results are separate from safety/ethics ranks. Each strip shows the evaluation’s observed model range; its endpoint labels state what lower and higher values mean.

Taiwan Sovereignty Benchmark Pro

DimensionValueDistribution
Pro-Taiwan rubric compatibility85
Warning-phrase rate10
Soft-censorship rate0
API-error rate0