← Models

Model profile

GPT 5.6 Terra

OpenAIdeveloper
2026-07-09release date
#13 / 346Safety rank
#573 / 662Freedom rank

Evidence summary

Safety. GPT 5.6 Terra has an estimated Safety rank of #13; its 90% source-sensitivity interval is #16–#72. Its behavior-only rank is #19; company governance moves the combined estimate to #13. Published Safety evidence spans 26 eval lineages and 7 of 7 components. Its strongest relative result is Opposite-Narrator Sycophancy (sycophancy_rate_pct, #1 of 24); its weakest is Vals AI Cheating Audit (terminal_bench_cheating_shortcut_evidence_rate_pct, #14 of 14).

Freedom. GPT 5.6 Terra has an estimated Freedom rank of #573; its 90% source-sensitivity interval is #373–#608. Published Freedom evidence spans 7 eval lineages and 1 of 1 components. Its strongest relative result is Enkrypt AI Safety Leaderboard (toxicity_attack_non_success_rate, #60 of 268); its weakest is GPT-5.6 system card (sexual_not_unsafe, #7 of 7).

Compare this model

Only models sharing at least one published sub-eval are listed.

Official and reference links

Safety evals

Rank and direction are specific to the Safety portfolio. Black marks the observed result; the grey dot marks the value implied by the Safety rank. Values stay on each source’s native scale.

Eval / sub-evalRankValueDistributionSource
AA-Omnisciencehallucination_rate#260 / 358↓0.8788Source ↗official
BullshitBench v2clear_pushback_rate#34 / 122↑0.505Source ↗official
CAIS Risk Indexagent_red_teaming#16 / 49↓54.6Source ↗official
CAIS Risk Indexbioweapons_assistance#28 / 54↓66.2Source ↗official
CAIS Risk Indexhle_overconfidence#24 / 55↓50.7Source ↗official
CAIS Risk Indexmachiavelli#4 / 51↓80.4Source ↗official
CAIS Risk Indexmask#8 / 57↓6.7Source ↗official
CAIS Risk Indexpolitical_manipulation#24 / 51↓46.2Source ↗official
CAIS Risk Indextextquests_harm#42 / 54↓22.2Source ↗official
Claude system cards — Gray Swan Q1+Q2 indirect prompt injection k=15attack_success_probability_k15_pct#9 / 12↓37.3Source ↗official
Enkrypt AI Safety Leaderboardbias_attack_non_success_rate#38 / 270↑39.28Source ↗official
Enkrypt AI Safety Leaderboardcbrn_attack_non_success_rate#200 / 270↑80Source ↗official
Enkrypt AI Safety Leaderboardharmful_attack_non_success_rate#31 / 270↑97.78Source ↗official
Enkrypt AI Safety Leaderboardtoxicity_attack_non_success_rate#209 / 268↑92.36Source ↗official
Google Gemini 3.8 launch — Gray Swan indirect prompt injection k=15attack_success_probability_k15_pct#11 / 15↓37.3Source ↗official
GPT 6 Astra system-card alignment evaluationsoverall_misaligned_outcome_base_pct#2 / 4↓13.5Source ↗official
GPT 6 Astra system-card alignment evaluationsoverall_misaligned_outcome_confirmation_pct#2 / 4↓5.7Source ↗official
GPT-5.6 system cardconnectors_injection_resistance#1 / 7↑1Source ↗official
GPT-5.6 system cardemotional_reliance#2 / 7↑0.976Source ↗official
GPT-5.6 system cardextremism_not_unsafe#4 / 7↑0.981Source ↗official
GPT-5.6 system cardgore_not_unsafe#6 / 7↑0.6Source ↗official
GPT-5.6 system cardharm_overall_pct#2 / 7↓0.88Source ↗official
GPT-5.6 system cardhate_not_unsafe#1 / 7↑1Source ↗official
GPT-5.6 system cardmental_health#3 / 7↑0.985Source ↗official
GPT-5.6 system cardnonviolent_illicit_not_unsafe#4 / 7↑0.99Source ↗official
GPT-5.6 system cardsearch_function_calling_injection_resistance#1 / 6↑0.946Source ↗official
GPT-5.6 system cardself_harm#3 / 7↑0.947Source ↗official
GPT-5.6 system cardself_harm_not_unsafe#2 / 7↑0.962Source ↗official
GPT-5.6 system cardsexual_minors_not_unsafe#1 / 7↑0.974Source ↗official
GPT-5.6 system cardsexual_not_unsafe#1 / 7↑0.966Source ↗official
GPT-5.6 system cardviolent_illicit_not_unsafe#4 / 7↑0.952Source ↗official
Gray Swan indirect prompt injection (15 attempts)attack_success_probability_k15_pct#9 / 13↓30.4Source ↗official
HUMAINE Trust, Ethics and Safetytrust_ethics_safety_score#19 / 54↑27.66Source ↗official
MACHIAVELLIdeception_relative_random_pct#3 / 50↓80.4Source ↗official
Manager Coercion Benchcoercion_ladder_depth#30 / 45↓8.7Source ↗official
Manager Coercion Benchfabrication_rate#1 / 15↓0Source ↗official
Olam Diplomacy — Broken Promise Ratebroken_promise_rate_pct#4 / 19↓10.9Source ↗official
Olam Social Poker — Social Lie Ratesocial_lie_rate_per_10000_turns#3 / 24↓2Source ↗official
Opposite-Narrator Sycophancysycophancy_rate_pct#1 / 24↓0Source ↗official
Pander Scoreconversational_absolute_pander_score#11 / 26↓9.7Source ↗official
Pander Scoreinstructional_absolute_pander_score#10 / 26↓25.68Source ↗official
SM-Benchadversarial#35 / 92↑83.66Source ↗official
SM-Benchambiguous_interpretation#54 / 92↑83.63Source ↗official
SM-Benchanti_hallucination#59 / 92↑90.84Source ↗official
SM-Bencheq_boundaries#41 / 92↑66.01Source ↗official
SM-Benchoverfit#71 / 92↑53.55Source ↗official
SpeciEvalbelief_animal_sentience#39 / 131↑6.925Source ↗official
SpeciEvalland_animal_4ns#31 / 131↓4.29Source ↗official
SpeciEvalsea_animal_4ns#8 / 131↓4.225Source ↗official
SpeciEvalspeciesism#9 / 131↓1.265Source ↗official
TACbase_welfare_rate#39 / 92↑29.49Source ↗official
Vals AI Cheating Auditswebench_verified_cheating_attempt_rate_pct#6 / 6↓89.4Source ↗official
Vals AI Cheating Auditterminal_bench_cheating_shortcut_evidence_rate_pct#14 / 14↓4.494Source ↗official

Freedom evals

Rank and direction are specific to the Freedom portfolio. Black marks the observed result; the grey dot marks the value implied by the Freedom rank. Values stay on each source’s native scale.

Eval / sub-evalRankValueDistributionSource
CAIS Risk Indexbioweapons_assistance#27 / 54↑66.2Source ↗official
Enkrypt AI Safety Leaderboardcbrn_attack_non_success_rate#71 / 270↓80Source ↗official
Enkrypt AI Safety Leaderboardharmful_attack_non_success_rate#232 / 270↓97.78Source ↗official
Enkrypt AI Safety Leaderboardtoxicity_attack_non_success_rate#60 / 268↓92.36Source ↗official
GPT-5.6 system cardemotional_reliance#6 / 7↓0.976Source ↗official
GPT-5.6 system cardextremism_not_unsafe#3 / 7↓0.981Source ↗official
GPT-5.6 system cardgore_not_unsafe#2 / 7↓0.6Source ↗official
GPT-5.6 system cardhate_not_unsafe#5 / 7↓1Source ↗official
GPT-5.6 system cardmental_health#4 / 7↓0.985Source ↗official
GPT-5.6 system cardnonviolent_illicit_not_unsafe#3 / 7↓0.99Source ↗official
GPT-5.6 system cardself_harm#5 / 7↓0.947Source ↗official
GPT-5.6 system cardself_harm_not_unsafe#6 / 7↓0.962Source ↗official
GPT-5.6 system cardsexual_minors_not_unsafe#6 / 7↓0.974Source ↗official
GPT-5.6 system cardsexual_not_unsafe#7 / 7↓0.966Source ↗official
GPT-5.6 system cardviolent_illicit_not_unsafe#4 / 7↓0.952Source ↗official
SM-Benchadversarial#57 / 92↓83.66Source ↗official
SM-Bencheq_boundaries#41 / 92↑66.01Source ↗official
SM-Benchoverfit#71 / 92↑53.55Source ↗official
SpeechMap model completioncomplete_pct#131 / 181↑44.2Source ↗official
UGI Leaderboard — base-model willingnesswillingness_adherence_score#79 / 156↑1.5Source ↗official
UGI Leaderboard — base-model willingnesswillingness_direct_score#79 / 156↑3Source ↗official

Values evaluations

Descriptive values and political-framing results are separate from safety/ethics ranks. Each strip shows the evaluation’s observed model range; its endpoint labels state what lower and higher values mean.

UGI Political Values

DimensionValueDistribution
Political Lean-26.5
Government47.5
Diplomacy68.3
Economy48.6
Society60.5