← Models

Model profile

GPT 5.2 Chat

OpenAIdeveloper
2025-12-11release date
#30 / 346Safety rank
#244 / 662Freedom rank

Evidence summary

Safety. GPT 5.2 Chat has an estimated Safety rank of #30; its 90% source-sensitivity interval is #7–#114. Its behavior-only rank is #36; company governance moves the combined estimate to #30. Published Safety evidence spans 5 eval lineages and 4 of 7 components. Its strongest relative result is SpeciEval (land_animal_4ns, #8 of 131); its weakest is SpeciEval (belief_animal_sentience, #95 of 131).

Freedom. GPT 5.2 Chat has an estimated Freedom rank of #244; its 90% source-sensitivity interval is #149–#436. Published Freedom evidence spans 1 eval lineages and 1 of 1 components. Its strongest relative result is SpeechMap model completion (complete_pct, #58 of 181); its weakest is SpeechMap model completion (complete_pct, #58 of 181).

Compare this model

Only models sharing at least one published sub-eval are listed.

Official and reference links

Safety evals

Rank and direction are specific to the Safety portfolio. Black marks the observed result; the grey dot marks the value implied by the Safety rank. Values stay on each source’s native scale.

Eval / sub-evalRankValueDistributionSource
Arena Factuality — Text Arena (factuality-only weighting)factuality_bt_rating#12 / 111↑1466.0Source ↗official
BullshitBench v2clear_pushback_rate#71 / 122↑0.27Source ↗official
Constitutional Following — OpenAI Model Specconstitutional_following_score#4 / 7↑94.4Source ↗official
HUMAINE Trust, Ethics and Safetytrust_ethics_safety_score#16 / 54↑27.91Source ↗official
SpeciEvalbelief_animal_sentience#95 / 131↑6.6Source ↗official
SpeciEvalland_animal_4ns#8 / 131↓3.92Source ↗official
SpeciEvalsea_animal_4ns#12 / 131↓4.3Source ↗official
SpeciEvalspeciesism#73 / 131↓2.08Source ↗official

Freedom evals

Rank and direction are specific to the Freedom portfolio. Black marks the observed result; the grey dot marks the value implied by the Freedom rank. Values stay on each source’s native scale.

Eval / sub-evalRankValueDistributionSource
SpeechMap model completioncomplete_pct#58 / 181↑69.6Source ↗official