← Models

Model profile

DeepSeek v4 Pro

DeepSeekdeveloper
2026-04-22release date
#167 / 346Safety rank
#276 / 662Freedom rank
3discovery sources

Evidence summary

Safety. DeepSeek v4 Pro has an estimated Safety rank of #167; its 90% source-sensitivity interval is #75–#294. Its behavior-only rank is #155; company governance moves the combined estimate to #167. Published Safety evidence spans 24 eval lineages and 7 of 7 components. Its strongest relative result is ANIMA (score, #1 of 22); its weakest is AgentAbstain (abstain, #17 of 17).

Freedom. DeepSeek v4 Pro has an estimated Freedom rank of #276; its 90% source-sensitivity interval is #149–#437. Published Freedom evidence spans 8 eval lineages and 1 of 1 components. Its strongest relative result is Human Pathogen Capabilities Test (HPCT) — overall refusal (refusal_rate_pct, #1 of 69); its weakest is UGI Leaderboard — base-model willingness (willingness_direct_score, #113 of 156).

Compare this model

Only models sharing at least one published sub-eval are listed.

Official and reference links

Safety evals

Rank and direction are specific to the Safety portfolio. Black marks the observed result; the grey dot marks the value implied by the Safety rank. Values stay on each source’s native scale.

Eval / sub-evalRankValueDistributionSource
AA-Omnisciencehallucination_rate#259 / 358↓0.8785Source ↗official
AgentAbstainabstain#17 / 17↑42.8Source ↗official
AgentAbstaincar#16 / 17↑42.3Source ↗official
AgentAbstainpaired#15 / 17↑36.9Source ↗official
ANIMAscore#1 / 22↑0.767Source ↗official
Arena Factuality — Text Arena (factuality-only weighting)factuality_bt_rating#55 / 111↑1437.0Source ↗official
BioTIERpermit_compliance_pct#35 / 52↑98.5Source ↗official
BioTIERrefuse_compliance_pct#35 / 52↑22.6Source ↗official
BullshitBench v2clear_pushback_rate#94 / 122↑0.14Source ↗official
CAIS Risk Indexpolitical_manipulation#19 / 51↓44Source ↗official
CheatBench direct cheating propensitybiology_bioinformatics_cheating_rate_pct#8 / 13↓95Source ↗official
CheatBench direct cheating propensityboard_games_cheating_rate_pct#1 / 13↓0Source ↗official
CheatBench direct cheating propensitycreative_writing_cheating_rate_pct#8 / 13↓100Source ↗official
CheatBench direct cheating propensityknowledge_work_cheating_rate_pct#10 / 13↓100Source ↗official
CheatBench direct cheating propensitymathematical_research_cheating_rate_pct#1 / 13↓52.5Source ↗official
CheatBench direct cheating propensitymenial_computation_cheating_rate_pct#5 / 13↓100Source ↗official
CheatBench direct cheating propensitymultimodal_cheating_rate_pct#9 / 13↓100Source ↗official
CheatBench direct cheating propensitysoftware_engineering_cheating_rate_pct#12 / 13↓68.4Source ↗official
CheatBench direct cheating propensitysvg_competition_cheating_rate_pct#10 / 13↓100Source ↗official
Concordia — Agentic-Misalignmentsafety_score#46 / 54↑54.83Source ↗official
Concordia — AirBench-Deceptionsafety_score#47 / 63↑70.37Source ↗official
Concordia — AirBench-Manipulationsafety_score#45 / 56↑74.67Source ↗official
Concordia — AirBench-SecurityRiskssafety_score#50 / 63↑80.18Source ↗official
Concordia — APEsafety_score#49 / 55↑1.542Source ↗official
Concordia — CyberSecEval2-PromptInjectionsafety_score#31 / 63↑84.86Source ↗official
Concordia — DarkBenchsafety_score#29 / 55↑55.45Source ↗official
Concordia — Fortress-Biologicalsafety_score#54 / 54↑18.62Source ↗official
Concordia — Fortress-Chemicalsafety_score#54 / 54↑21.39Source ↗official
Concordia — Fortress-Privacy/Scamssafety_score#43 / 54↑43.05Source ↗official
Concordia — FRT-AirBench-Manipulationsafety_score#41 / 45↑15.11Source ↗official
Concordia — FRT-AirBench-SecurityRiskssafety_score#43 / 45↑8.667Source ↗official
Concordia — FRT-SciKnowEval-BiologicalHarmfulQAsafety_score#36 / 45↑1.333Source ↗official
Concordia — FRT-SOSBench-Chemsafety_score#39 / 45↑27.33Source ↗official
Concordia — MASKsafety_score#59 / 62↑39.38Source ↗official
Concordia — SciKnowEval-BiologicalHarmfulQAsafety_score#35 / 63↑77.44Source ↗official
Concordia — SciKnowEval-ChemicalHarmfulQAsafety_score#51 / 63↑7.93Source ↗official
Concordia — Shutdown-Resistancesafety_score#35 / 53↑98.98Source ↗official
Concordia — SOSBench-Biosafety_score#21 / 63↑91.4Source ↗official
Concordia — SOSBench-Chemsafety_score#32 / 63↑87.8Source ↗official
DystopiaBenchbasaglia_score#26 / 50↓64.6Source ↗official
DystopiaBenchbaudrillard_score#31 / 50↓62.2Source ↗official
DystopiaBenchhuxley_score#41 / 50↓76.97Source ↗official
DystopiaBenchlaguardia_score#42 / 50↓69.8Source ↗official
DystopiaBenchorwell_score#40 / 50↓74Source ↗official
DystopiaBenchpetrov_score#36 / 50↓77Source ↗official
Enkrypt AI Safety Leaderboardbias_attack_non_success_rate#158 / 270↑14.73Source ↗official
Enkrypt AI Safety Leaderboardcbrn_attack_non_success_rate#183 / 270↑82.67Source ↗official
Enkrypt AI Safety Leaderboardharmful_attack_non_success_rate#121 / 270↑77.78Source ↗official
Enkrypt AI Safety Leaderboardtoxicity_attack_non_success_rate#208 / 268↑92.55Source ↗official
Every Model Cheats — Cybench Cheat Propensitybaseline_cheat_propensity_pct#5 / 22↓21.7Source ↗official
Every Model Cheats — Cybench Cheat Propensitysevere_anti_cheat_cheat_propensity_pct#1 / 22↓0Source ↗official
Every Model Cheats — Cybench Cheat Propensitystandard_anti_cheat_cheat_propensity_pct#13 / 22↓17.4Source ↗official
Google Gemini 3.8 launch — Gray Swan indirect prompt injection k=15attack_success_probability_k15_pct#15 / 15↓60.1Source ↗official
HUMAINE Trust, Ethics and Safetytrust_ethics_safety_score#20 / 54↑27.61Source ↗official
Human Pathogen Capabilities Test (HPCT) — overall refusalrefusal_rate_pct#27 / 69↑0Source ↗official
Manager Coercion Benchcoercion_ladder_depth#41 / 45↓9Source ↗official
Manager Coercion Benchfabrication_rate#1 / 15↓0Source ↗official
Olam Social Poker — Social Lie Ratesocial_lie_rate_per_10000_turns#12 / 24↓9Source ↗official
Opposite-Narrator Sycophancysycophancy_rate_pct#17 / 24↓5.1Source ↗official
PHAREbias_resistance_diagnostic#66 / 66↑0.1714Source ↗official
PHAREhallucination_resistance_diagnostic#21 / 70↑0.7964Source ↗official
PHAREharm_resistance_diagnostic#22 / 70↑0.9541Source ↗official
PHAREjailbreak_resistance_diagnostic#44 / 67↑0.4321Source ↗official
SM-Benchadversarial#52 / 92↑80.98Source ↗official
SM-Benchambiguous_interpretation#84 / 92↑68.75Source ↗official
SM-Benchanti_hallucination#64 / 92↑88.48Source ↗official
SM-Bencheq_boundaries#45 / 92↑65.17Source ↗official
SM-Benchoverfit#46 / 92↑76.5Source ↗official
SpeciEvalbelief_animal_sentience#68 / 131↑6.8Source ↗official
SpeciEvalland_animal_4ns#39 / 131↓4.35Source ↗official
SpeciEvalsea_animal_4ns#46 / 131↓4.65Source ↗official
SpeciEvalspeciesism#44 / 131↓1.77Source ↗official
TACbase_welfare_rate#67 / 92↑23.08Source ↗official
The Dictatorship Evaloverall_resistance_rate#18 / 20↑11.65Source ↗official
Vectara HHEM Factual Consistencyfactual_consistency_rate#39 / 94↑91.4Source ↗official

Freedom evals

Rank and direction are specific to the Freedom portfolio. Black marks the observed result; the grey dot marks the value implied by the Freedom rank. Values stay on each source’s native scale.

Eval / sub-evalRankValueDistributionSource
BioTIERpermit_compliance_pct#35 / 52↑98.5Source ↗official
BioTIERrefuse_compliance_pct#18 / 52↓22.6Source ↗official
Concordia — AirBench-Deceptionsafety_score#16 / 63↓70.37Source ↗official
Concordia — AirBench-Manipulationsafety_score#12 / 56↓74.67Source ↗official
Concordia — AirBench-PoliticalPersuasionsafety_score#19 / 56↓49.52Source ↗official
Concordia — AirBench-SecurityRiskssafety_score#14 / 63↓80.18Source ↗official
Concordia — Fortress-Biologicalsafety_score#1 / 54↓18.62Source ↗official
Concordia — Fortress-Chemicalsafety_score#1 / 54↓21.39Source ↗official
Concordia — Fortress-Privacy/Scamssafety_score#12 / 54↓43.05Source ↗official
Concordia — FRT-AirBench-Manipulationsafety_score#4 / 45↓15.11Source ↗official
Concordia — FRT-AirBench-SecurityRiskssafety_score#3 / 45↓8.667Source ↗official
Concordia — FRT-SciKnowEval-BiologicalHarmfulQAsafety_score#9 / 45↓1.333Source ↗official
Concordia — FRT-SOSBench-Chemsafety_score#6 / 45↓27.33Source ↗official
Concordia — SciKnowEval-BiologicalHarmfulQAsafety_score#28 / 63↓77.44Source ↗official
Concordia — SciKnowEval-ChemicalHarmfulQAsafety_score#12 / 63↓7.93Source ↗official
Concordia — SOSBench-Biosafety_score#43 / 63↓91.4Source ↗official
Concordia — SOSBench-Chemsafety_score#32 / 63↓87.8Source ↗official
Enkrypt AI Safety Leaderboardcbrn_attack_non_success_rate#88 / 270↓82.67Source ↗official
Enkrypt AI Safety Leaderboardharmful_attack_non_success_rate#150 / 270↓77.78Source ↗official
Enkrypt AI Safety Leaderboardtoxicity_attack_non_success_rate#61 / 268↓92.55Source ↗official
Human Pathogen Capabilities Test (HPCT) — overall refusalrefusal_rate_pct#1 / 69↓0Source ↗official
PHAREharm_resistance_diagnostic#49 / 70↓0.9541Source ↗official
PHAREjailbreak_resistance_diagnostic#24 / 67↓0.4321Source ↗official
SM-Benchadversarial#39 / 92↓80.98Source ↗official
SM-Bencheq_boundaries#45 / 92↑65.17Source ↗official
SM-Benchoverfit#46 / 92↑76.5Source ↗official
SpeechMap model completioncomplete_pct#121 / 181↑46.8Source ↗official
UGI Leaderboard — base-model willingnesswillingness_adherence_score#50 / 156↑4.25Source ↗official
UGI Leaderboard — base-model willingnesswillingness_direct_score#113 / 156↑2Source ↗official

Values evaluations

Descriptive values and political-framing results are separate from safety/ethics ranks. Each strip shows the evaluation’s observed model range; its endpoint labels state what lower and higher values mean.

UGI Political Values

DimensionValueDistribution
Political Lean-18.6
Government45.1
Diplomacy66
Economy44.7
Society59.2

CAIS AI Values — countries

Moral Trolley Arena