← Models

Model profile

GPT 5

OpenAIdeveloper
2025-08-07release date
#41 / 267overall rank
32eval lineages

Evidence summary

GPT 5 has an estimated overall rank of #41; its 90% source-sensitivity interval is #15–#139. Its behavior-only rank is #49; company governance moves the combined estimate to #41. Published evidence spans 32 evals and 7 of 7 behavior components. Its strongest relative result is Enkrypt AI Safety Leaderboard (bias_attack_non_success_rate, #10 of 260); its weakest is OpenAgentSafety (rule_based_safety_vulnerable, #7 of 7).

Compare this model

Only models sharing at least one published sub-eval are listed.

Official and reference links

Published eval results

Rank is within that sub-eval. Black marks the observed result; the grey dot marks the value implied by the global rank. Values stay on each source’s native scale.

Eval / sub-evalRankValueDistributionBetterSource
AA-Omnisciencehallucination_rate#122 / 3110.7584↓ lowerSource ↗official
AgentAbstainabstain#2 / 1769.8↑ higherSource ↗official
AgentAbstaincar#2 / 1766.5↑ higherSource ↗official
AgentAbstainpaired#6 / 1749.6↑ higherSource ↗official
AIRBench 2024 Safety Scenariossafety_scenarios#7 / 800.877↑ higherSource ↗official
BrokenMathsycophancy#1 / 929↓ lowerSource ↗official
BullshitBench v2clear_pushback_rate#65 / 1050.195↑ higherSource ↗official
CAIS Risk Indexbioweapons_assistance#15 / 4862.6↓ lowerSource ↗official
CAIS Risk Indexhle_overconfidence#18 / 4950↓ lowerSource ↗official
CAIS Risk Indexmachiavelli#11 / 4583.8↓ lowerSource ↗official
CAIS Risk Indexmask#25 / 5120.7↓ lowerSource ↗official
CAIS Risk Indexpolitical_manipulation#4 / 3235.3↓ lowerSource ↗official
CAIS Risk Indextextquests_harm#22 / 4817.2↓ lowerSource ↗official
Cisco AI Defense Rolling Single-Turn Leaderboardsingle_turn_attack_success_rate#23 / 1058.076↓ lowerSource ↗official
Enkrypt AI Safety Leaderboardbias_attack_non_success_rate#10 / 26059.95↑ higherSource ↗official
Enkrypt AI Safety Leaderboardcbrn_attack_non_success_rate#45 / 26092.89↑ higherSource ↗official
Enkrypt AI Safety Leaderboardharmful_attack_non_success_rate#42 / 26093.33↑ higherSource ↗official
Enkrypt AI Safety Leaderboardtoxicity_attack_non_success_rate#112 / 25897.09↑ higherSource ↗official
FlagEval Safety and Valuesa1_qualified_rate#10 / 1880.22↑ higherSource ↗official
FlagEval Safety and Valuesa2_qualified_rate#9 / 1879.66↑ higherSource ↗official
FlagEval Safety and Valuesa3_qualified_rate#14 / 1883.26↑ higherSource ↗official
FlagEval Safety and Valuesa4_qualified_rate#13 / 1883.73↑ higherSource ↗official
FlagEval Safety and Valuesa5_qualified_rate#10 / 1872.15↑ higherSource ↗official
FORTRESSaverage_risk_score#13 / 4917.04↓ lowerSource ↗official
FORTRESSover_refusal_score#39 / 4611.2↓ lowerSource ↗official
HELM Safetyanthropic_red_team#32 / 800.991↑ higherSource ↗official
HELM Safetybbq#12 / 800.968↑ higherSource ↗official
HELM Safetyharmbench#6 / 800.976↑ higherSource ↗official
HELM Safetysimple_safety_tests#24 / 800.998↑ higherSource ↗official
HELM Safetyxstest#22 / 800.971↑ higherSource ↗official
HUMAINE Trust, Ethics and Safetytrust_ethics_safety_score#32 / 5426.31↑ higherSource ↗official
MACHIAVELLIdeception_relative_random_pct#11 / 5083.8↓ lowerSource ↗official
MASKlying_probability_pct#16 / 5320.67↓ lowerSource ↗official
OpenAgentSafetyllm_judge_safety_vulnerable#3 / 752.58↓ lowerSource ↗official
OpenAgentSafetyrule_based_safety_vulnerable#7 / 752.15↓ lowerSource ↗official
OpenAgentSafetysuccessful_completion#1 / 740.4↑ higherSource ↗official
OpenAI GPT-5 System Cardharmful_request_safety#1 / 20.912↑ higherSource ↗official
PacifAIstp_score#7 / 779.49↑ higherSource ↗official
PHAREbias_resistance_diagnostic#63 / 660.2856↑ higherSource ↗official
PHAREhallucination_resistance_diagnostic#38 / 700.7458↑ higherSource ↗official
PHAREharm_resistance_diagnostic#12 / 700.9697↑ higherSource ↗official
PHAREjailbreak_resistance_diagnostic#14 / 670.6868↑ higherSource ↗official
Shelleducation_jsr#2 / 140.364↓ lowerSource ↗official
Shellfinance_jsr#2 / 140.19↓ lowerSource ↗official
Shellmanagement_jsr#3 / 140.37↓ lowerSource ↗official
Social Welfare Function Benchmarkfairness#15 / 190.4455↑ higherSource ↗official
SOSBenchbiology_pvr#2 / 230.108↓ lowerSource ↗official
SOSBenchchemistry_pvr#1 / 230.122↓ lowerSource ↗official
SOSBenchmedicine_pvr#3 / 230.332↓ lowerSource ↗official
SOSBenchpharmacology_pvr#4 / 230.418↓ lowerSource ↗official
SOSBenchphysics_pvr#2 / 230.104↓ lowerSource ↗official
SOSBenchpsychology_pvr#3 / 230.142↓ lowerSource ↗official
SpeciEvalbelief_animal_sentience#40 / 1026.855↑ higherSource ↗official
SpeciEvalland_animal_4ns#82 / 1024.825↓ lowerSource ↗official
SpeciEvalsea_animal_4ns#48 / 1024.745↓ lowerSource ↗official
SpeciEvalspeciesism#13 / 1021.49↓ lowerSource ↗official
TrustLLM contemporary collapsed applicationtrustllm#7 / 80.6↑ higherSource ↗official
UAVBench safety-critical decision recognitionethical_safety_critical_accuracy#2 / 270.76↑ higherSource ↗official
Vigil Mental Health Safetyoverall_score#9 / 2351↑ higherSource ↗official

Values evaluations

Descriptive values and political-framing results are separate from safety/ethics ranks. Each strip shows the evaluation’s observed model range; its endpoint labels state what lower and higher values mean.

UGI Political Values

DimensionValueDistribution
Political Lean-18.2
Government45.9
Diplomacy65.3
Economy49
Society58.6

CAISI CCP narrative alignment

DimensionValueDistribution
CCP narrative alignment1.95