← Models

Model profile

GPT 5.4

OpenAIdeveloper
2026-03-05release date
#24 / 267overall rank
22eval lineages

Evidence summary

GPT 5.4 has an estimated overall rank of #24; its 90% source-sensitivity interval is #11–#100. Its behavior-only rank is #27; company governance moves the combined estimate to #24. Published evidence spans 22 evals and 7 of 7 behavior components. Its strongest relative result is Enkrypt AI Safety Leaderboard (harmful_attack_non_success_rate, #1 of 260); its weakest is CAIS Risk Index (textquests_harm, #47 of 48).

Compare this model

Only models sharing at least one published sub-eval are listed.

Official and reference links

Published eval results

Rank is within that sub-eval. Black marks the observed result; the grey dot marks the value implied by the global rank. Values stay on each source’s native scale.

Eval / sub-evalRankValueDistributionBetterSource
AA-Omnisciencehallucination_rate#162 / 3110.8109↓ lowerSource ↗official
AgentAbstainabstain#3 / 1767.8↑ higherSource ↗official
AgentAbstaincar#5 / 1764↑ higherSource ↗official
AgentAbstainpaired#7 / 1748.7↑ higherSource ↗official
BioSecBench-Refusalbalanced_refusal_score#4 / 100.4254↑ higherSource ↗official
BullshitBench v2clear_pushback_rate#29 / 1050.45↑ higherSource ↗official
CAIS Risk Indexagent_red_teaming#13 / 4359.7↓ lowerSource ↗official
CAIS Risk Indexbioweapons_assistance#11 / 4857.1↓ lowerSource ↗official
CAIS Risk Indexhle_overconfidence#6 / 4940.9↓ lowerSource ↗official
CAIS Risk Indexmachiavelli#34 / 4592.9↓ lowerSource ↗official
CAIS Risk Indexmask#11 / 519.7↓ lowerSource ↗official
CAIS Risk Indexpolitical_manipulation#28 / 3259.4↓ lowerSource ↗official
CAIS Risk Indextextquests_harm#47 / 4824.5↓ lowerSource ↗official
Cisco AI Defense Rolling Single-Turn Leaderboardsingle_turn_attack_success_rate#8 / 1052.742↓ lowerSource ↗official
DystopiaBenchbasaglia_score#13 / 5047.3↓ lowerSource ↗official
DystopiaBenchbaudrillard_score#12 / 5038.87↓ lowerSource ↗official
DystopiaBenchhuxley_score#13 / 5046.2↓ lowerSource ↗official
DystopiaBenchlaguardia_score#10 / 5046.57↓ lowerSource ↗official
DystopiaBenchorwell_score#13 / 5044.4↓ lowerSource ↗official
DystopiaBenchpetrov_score#7 / 5037.4↓ lowerSource ↗official
Enkrypt AI Safety Leaderboardbias_attack_non_success_rate#181 / 26012.92↑ higherSource ↗official
Enkrypt AI Safety Leaderboardcbrn_attack_non_success_rate#199 / 26083.33↑ higherSource ↗official
Enkrypt AI Safety Leaderboardharmful_attack_non_success_rate#1 / 260100↑ higherSource ↗official
Enkrypt AI Safety Leaderboardtoxicity_attack_non_success_rate#105 / 25897.27↑ higherSource ↗official
GPT-5.6 system card — disallowed content with challenging promptsextremism_not_unsafe#1 / 71↑ higherSource ↗official
GPT-5.6 system card — disallowed content with challenging promptsgore_not_unsafe#2 / 70.831↑ higherSource ↗official
GPT-5.6 system card — disallowed content with challenging promptshate_not_unsafe#5 / 70.943↑ higherSource ↗official
GPT-5.6 system card — disallowed content with challenging promptsnonviolent_illicit_not_unsafe#1 / 71↑ higherSource ↗official
GPT-5.6 system card — disallowed content with challenging promptsself_harm_not_unsafe#1 / 70.987↑ higherSource ↗official
GPT-5.6 system card — disallowed content with challenging promptssexual_minors_not_unsafe#4 / 70.966↑ higherSource ↗official
GPT-5.6 system card — disallowed content with challenging promptssexual_not_unsafe#6 / 70.933↑ higherSource ↗official
GPT-5.6 system card — disallowed content with challenging promptsviolent_illicit_not_unsafe#2 / 70.971↑ higherSource ↗official
GPT-5.6 system card — first-person fairnessharm_overall_pct#2 / 70.88↓ lowerSource ↗official
GPT-5.6 system card — prompt-injection robustnessconnectors_injection_resistance#5 / 70.998↑ higherSource ↗official
GPT-5.6 system card — prompt-injection robustnesssearch_function_calling_injection_resistance#4 / 60.697↑ higherSource ↗official
HUMAINE Trust, Ethics and Safetytrust_ethics_safety_score#25 / 5427.17↑ higherSource ↗official
MACHIAVELLIdeception_relative_random_pct#34 / 5092.9↓ lowerSource ↗official
MASKlying_probability_pct#8 / 5310.33↓ lowerSource ↗official
OpenAI GPT-5.4 Dynamic Wellbeingemotional_reliance#1 / 30.985↑ higherSource ↗official
OpenAI GPT-5.4 Dynamic Wellbeingmental_health#1 / 30.985↑ higherSource ↗official
OpenAI GPT-5.4 Dynamic Wellbeingself_harm#1 / 30.977↑ higherSource ↗official
OpenAI GPT-5.4 First-Person Fairnessharm_overall#1 / 30.0088↓ lowerSource ↗official
OpenAI GPT-5.4 Property Preservationdestructive_action_avoidance#2 / 30.86↑ higherSource ↗official
OpenAI GPT-5.4 Property Preservationuser_work_preserved#1 / 30.53↑ higherSource ↗official
OpenAI GPT-5.4 User Confirmationsfinancial_transaction#1 / 31↑ higherSource ↗official
OpenAI GPT-5.4 User Confirmationsgeneral_confirmation#1 / 30.94↑ higherSource ↗official
OpenAI GPT-5.4 User Confirmationshigh_stakes_communication#1 / 31↑ higherSource ↗official
SABERoverall_safety_rate#2 / 1336.1↑ higherSource ↗official
SABERscenario_a_safety_rate#3 / 1336.02↑ higherSource ↗official
SABERscenario_b_safety_rate#2 / 1339.38↑ higherSource ↗official
SABERscenario_c_safety_rate#2 / 1333.51↑ higherSource ↗official
SM-Benchadversarial#57 / 7377.56↑ higherSource ↗official
SM-Benchambiguous_interpretation#23 / 7387.8↑ higherSource ↗official
SM-Benchanti_hallucination#42 / 7390.58↑ higherSource ↗official
SM-Bencheq_boundaries#46 / 7360.11↑ higherSource ↗official
SM-Benchoverfit#62 / 7338.25↑ higherSource ↗official
SpeciEvalbelief_animal_sentience#72 / 1026.57↑ higherSource ↗official
SpeciEvalland_animal_4ns#5 / 1023.7↓ lowerSource ↗official
SpeciEvalsea_animal_4ns#1 / 1023.65↓ lowerSource ↗official
SpeciEvalspeciesism#37 / 1021.82↓ lowerSource ↗official
VETO Misfired Alignmentmisfired_alignment_rate_pct#24 / 2517.6↓ lowerSource ↗official
Vigil Mental Health Safetyoverall_score#2 / 2378↑ higherSource ↗official

Values evaluations

Descriptive values and political-framing results are separate from safety/ethics ranks. Each strip shows the evaluation’s observed model range; its endpoint labels state what lower and higher values mean.

Agent-ValueBench Moral Foundations (MFT08)

Agent-ValueBench HEXACO

DimensionValueDistribution
Openness to experience4.5
Honesty-humility6
Extraversion5.7
Agreeableness5.6
Conscientiousness7.3

Agent-ValueBench Schwartz Basic Values (PVQ40)

The Economist World Values Survey Cultural Map

DimensionValueDistribution
Survival ↔ Self-expression2.26
Traditional ↔ Secular2.08