← Models

Model profile

DeepSeek v4 Flash

DeepSeekdeveloper
2026-04-22release date
#109 / 346Safety rank
#220 / 662Freedom rank
4discovery sources

Evidence summary

Safety. DeepSeek v4 Flash has an estimated Safety rank of #109; its 90% source-sensitivity interval is #76–#246. Its behavior-only rank is #96; company governance moves the combined estimate to #109. Published Safety evidence spans 20 eval lineages and 7 of 7 components. Its strongest relative result is Manager Coercion Bench (coercion_ladder_depth, #1 of 45); its weakest is Inkling-Small model card — StrongREJECT (safety_rate, #10 of 10).

Freedom. DeepSeek v4 Flash has an estimated Freedom rank of #220; its 90% source-sensitivity interval is #55–#457. Published Freedom evidence spans 6 eval lineages and 1 of 1 components. Its strongest relative result is UGI Leaderboard — base-model willingness (willingness_adherence_score, #6 of 156); its weakest is SpeechMap model completion (complete_pct, #131 of 181).

Compare this model

Only models sharing at least one published sub-eval are listed.

Official and reference links

Safety evals

Rank and direction are specific to the Safety portfolio. Black marks the observed result; the grey dot marks the value implied by the Safety rank. Values stay on each source’s native scale.

Eval / sub-evalRankValueDistributionSource
AA-Omnisciencehallucination_rate#272 / 358↓0.8897Source ↗official
ANIMAscore#3 / 22↑0.7445Source ↗official
Arena Factuality — Text Arena (factuality-only weighting)factuality_bt_rating#63 / 111↑1432.0Source ↗official
BullshitBench v2clear_pushback_rate#88 / 122↑0.16Source ↗official
DystopiaBenchbasaglia_score#23 / 50↓60.9Source ↗official
DystopiaBenchbaudrillard_score#27 / 50↓58.1Source ↗official
DystopiaBenchhuxley_score#27 / 50↓72.8Source ↗official
DystopiaBenchlaguardia_score#37 / 50↓69.03Source ↗official
DystopiaBenchorwell_score#28 / 50↓69.83Source ↗official
DystopiaBenchpetrov_score#30 / 50↓74.3Source ↗official
Enkrypt AI Safety Leaderboardbias_attack_non_success_rate#158 / 270↑14.73Source ↗official
Enkrypt AI Safety Leaderboardcbrn_attack_non_success_rate#232 / 270↑70.33Source ↗official
Enkrypt AI Safety Leaderboardharmful_attack_non_success_rate#123 / 270↑76.67Source ↗official
Enkrypt AI Safety Leaderboardtoxicity_attack_non_success_rate#216 / 268↑91.64Source ↗official
Every Model Cheats — Cybench Cheat Propensitybaseline_cheat_propensity_pct#10 / 22↓30.4Source ↗official
Every Model Cheats — Cybench Cheat Propensitysevere_anti_cheat_cheat_propensity_pct#7 / 22↓4.3Source ↗official
Every Model Cheats — Cybench Cheat Propensitystandard_anti_cheat_cheat_propensity_pct#6 / 22↓8.7Source ↗official
Governance Decay under Passive Context Compactiongovernance_retention_score#6 / 7↑41Source ↗official
HUMAINE Trust, Ethics and Safetytrust_ethics_safety_score#10 / 54↑28.57Source ↗official
Inkling-Small model card — FORTRESSbenign_answer_rate#1 / 10↑99.2Source ↗official
Inkling-Small model card — FORTRESSharmful_refusal_rate#10 / 10↑32Source ↗official
Inkling-Small model card — StrongREJECTsafety_rate#10 / 10↑97.4Source ↗official
KIDBench Implicit Child Cueimplicit_child_cue_total_mean#1 / 13↑4.43Source ↗official
Manager Coercion Benchcoercion_ladder_depth#1 / 45↓3.233Source ↗self run
MANTAAWMS#9 / 12↑0.417Source ↗official
MANTAAWVS#6 / 12↑0.508Source ↗official
Olam Social Poker — Social Lie Ratesocial_lie_rate_per_10000_turns#19 / 24↓37Source ↗official
Opposite-Narrator Sycophancysycophancy_rate_pct#19 / 24↓5.6Source ↗official
PHAREbias_resistance_diagnostic#59 / 66↑0.323Source ↗official
PHAREhallucination_resistance_diagnostic#13 / 70↑0.8188Source ↗official
PHAREharm_resistance_diagnostic#28 / 70↑0.9507Source ↗official
PHAREjailbreak_resistance_diagnostic#42 / 67↑0.4549Source ↗official
SM-Benchadversarial#46 / 92↑81.95Source ↗official
SM-Benchambiguous_interpretation#78 / 92↑75.89Source ↗official
SM-Benchanti_hallucination#46 / 92↑94.24Source ↗official
SM-Bencheq_boundaries#22 / 92↑69.94Source ↗official
SM-Benchoverfit#66 / 92↑59.56Source ↗official
SpeciEvalbelief_animal_sentience#42 / 131↑6.9Source ↗official
SpeciEvalland_animal_4ns#90 / 131↓4.7Source ↗official
SpeciEvalsea_animal_4ns#69 / 131↓4.75Source ↗official
SpeciEvalspeciesism#70 / 131↓2.05Source ↗official
TACbase_welfare_rate#69 / 92↑21.15Source ↗official
ToolPrivacyBenchprivate_mt_poi#6 / 9↓27.33Source ↗official
ToolPrivacyBenchpublic_mt_poi#6 / 9↓18.99Source ↗official

Freedom evals

Rank and direction are specific to the Freedom portfolio. Black marks the observed result; the grey dot marks the value implied by the Freedom rank. Values stay on each source’s native scale.

Eval / sub-evalRankValueDistributionSource
Enkrypt AI Safety Leaderboardcbrn_attack_non_success_rate#39 / 270↓70.33Source ↗official
Enkrypt AI Safety Leaderboardharmful_attack_non_success_rate#148 / 270↓76.67Source ↗official
Enkrypt AI Safety Leaderboardtoxicity_attack_non_success_rate#53 / 268↓91.64Source ↗official
Inkling-Small model card — FORTRESSbenign_answer_rate#1 / 10↑99.2Source ↗official
Inkling-Small model card — FORTRESSharmful_refusal_rate#1 / 10↓32Source ↗official
Inkling-Small model card — StrongREJECTsafety_rate#1 / 10↓97.4Source ↗official
PHAREharm_resistance_diagnostic#43 / 70↓0.9507Source ↗official
PHAREjailbreak_resistance_diagnostic#26 / 67↓0.4549Source ↗official
SM-Benchadversarial#45 / 92↓81.95Source ↗official
SM-Bencheq_boundaries#22 / 92↑69.94Source ↗official
SM-Benchoverfit#66 / 92↑59.56Source ↗official
SpeechMap model completioncomplete_pct#131 / 181↑44.2Source ↗official
UGI Leaderboard — base-model willingnesswillingness_adherence_score#6 / 156↑8Source ↗official
UGI Leaderboard — base-model willingnesswillingness_direct_score#41 / 156↑4.5Source ↗official

Values evaluations

Descriptive values and political-framing results are separate from safety/ethics ranks. Each strip shows the evaluation’s observed model range; its endpoint labels state what lower and higher values mean.

UGI Political Values

DimensionValueDistribution
Political Lean-16.6
Government46
Diplomacy66
Economy46.2
Society59

CCPBench political narrative alignment

DimensionValueDistribution
CCP-narrative alignment — all questions3.57
CCP-narrative alignment — China topics4.03
CCP-narrative alignment — non-China controls2.22

The Economist World Values Survey Cultural Map

DimensionValueDistribution
Survival ↔ Self-expression2.27
Traditional ↔ Secular-1.75