← Models

Model profile

GPT 5.2 Chat

OpenAIdeveloper
2025-12-11release date
#13 / 267overall rank
3eval lineages

Evidence summary

GPT 5.2 Chat has an estimated overall rank of #13; its 90% source-sensitivity interval is #3–#132. Published evidence spans 3 evals and 3 of 7 behavior components. Its strongest relative result is SpeciEval (land_animal_4ns, #7 of 102); its weakest is SpeciEval (belief_animal_sentience, #70 of 102).

Compare this model

Only models sharing at least one published sub-eval are listed.

Official and reference links

Published eval results

Rank is within that sub-eval. Black marks the observed result; the grey dot marks the value implied by the global rank. Values stay on each source’s native scale.

Eval / sub-evalRankValueDistributionBetterSource
BullshitBench v2clear_pushback_rate#54 / 1050.27↑ higherSource ↗official
HUMAINE Trust, Ethics and Safetytrust_ethics_safety_score#16 / 5427.91↑ higherSource ↗official
SpeciEvalbelief_animal_sentience#70 / 1026.6↑ higherSource ↗official
SpeciEvalland_animal_4ns#7 / 1023.92↓ lowerSource ↗official
SpeciEvalsea_animal_4ns#9 / 1024.3↓ lowerSource ↗official
SpeciEvalspeciesism#55 / 1022.08↓ lowerSource ↗official