← Models

Model profile

Qwen3.5 Flash

Alibabadeveloper
2026-02-24release date
#189 / 346Safety rank
#586 / 662Freedom rank

Evidence summary

Safety. Qwen3.5 Flash has an estimated Safety rank of #189; its 90% source-sensitivity interval is #92–#303. Its behavior-only rank is #186; company governance moves the combined estimate to #189. Published Safety evidence spans 6 eval lineages and 5 of 7 components. Its strongest relative result is Adversarial Humanities Benchmark (AHB) — Table 5 (sexual_content, #4 of 31); its weakest is StereoTales Harmful Associations (benign_significant_association_score, #16 of 23).

Freedom. Qwen3.5 Flash has an estimated Freedom rank of #586; its 90% source-sensitivity interval is #411–#607. Published Freedom evidence spans 3 eval lineages and 1 of 1 components. Its strongest relative result is Adversarial Humanities Benchmark (AHB) — Table 5 (privacy, #22 of 31); its weakest is Adversarial Humanities Benchmark (AHB) — Table 5 (sexual_content, #27 of 31).

Compare this model

Only models sharing at least one published sub-eval are listed.

Official and reference links

Safety evals

Rank and direction are specific to the Safety portfolio. Black marks the observed result; the grey dot marks the value implied by the Safety rank. Values stay on each source’s native scale.

Eval / sub-evalRankValueDistributionSource
Adversarial Humanities Benchmark (AHB) — Table 5child_sexual_exploitation#8 / 31↓21.3Source ↗official
Adversarial Humanities Benchmark (AHB) — Table 5defamation#7 / 31↓44.4Source ↗official
Adversarial Humanities Benchmark (AHB) — Table 5hate#5 / 31↓30.4Source ↗official
Adversarial Humanities Benchmark (AHB) — Table 5indiscriminate_weapons_cbrne#9 / 31↓31.1Source ↗official
Adversarial Humanities Benchmark (AHB) — Table 5intellectual_property#9 / 31↓37.5Source ↗official
Adversarial Humanities Benchmark (AHB) — Table 5non_violent_crimes#9 / 31↓43.2Source ↗official
Adversarial Humanities Benchmark (AHB) — Table 5privacy#10 / 31↓48.9Source ↗official
Adversarial Humanities Benchmark (AHB) — Table 5sex_related_crimes#10 / 31↓41.7Source ↗official
Adversarial Humanities Benchmark (AHB) — Table 5sexual_content#4 / 31↓22.7Source ↗official
Adversarial Humanities Benchmark (AHB) — Table 5specialized_advice#8 / 31↓42.38Source ↗official
Adversarial Humanities Benchmark (AHB) — Table 5suicide_self_harm#8 / 31↓27.7Source ↗official
Adversarial Humanities Benchmark (AHB) — Table 5violent_crimes#7 / 31↓37Source ↗official
Arena Factuality — Text Arena (factuality-only weighting)factuality_bt_rating#72 / 111↑1428.0Source ↗official
MT-JailBench CrescendoXsafety_score#7 / 21↑34.59Source ↗official
StereoTales Harmful Associationsbenign_significant_association_score#16 / 23↑85.53Source ↗official
TACbase_welfare_rate#57 / 92↑24.36Source ↗self run
Vectara HHEM Factual Consistencyfactual_consistency_rate#56 / 94↑89.5Source ↗official

Freedom evals

Rank and direction are specific to the Freedom portfolio. Black marks the observed result; the grey dot marks the value implied by the Freedom rank. Values stay on each source’s native scale.

Eval / sub-evalRankValueDistributionSource
Adversarial Humanities Benchmark (AHB) — Table 5child_sexual_exploitation#24 / 31↑21.3Source ↗official
Adversarial Humanities Benchmark (AHB) — Table 5defamation#25 / 31↑44.4Source ↗official
Adversarial Humanities Benchmark (AHB) — Table 5hate#26 / 31↑30.4Source ↗official
Adversarial Humanities Benchmark (AHB) — Table 5indiscriminate_weapons_cbrne#23 / 31↑31.1Source ↗official
Adversarial Humanities Benchmark (AHB) — Table 5intellectual_property#23 / 31↑37.5Source ↗official
Adversarial Humanities Benchmark (AHB) — Table 5non_violent_crimes#23 / 31↑43.2Source ↗official
Adversarial Humanities Benchmark (AHB) — Table 5privacy#22 / 31↑48.9Source ↗official
Adversarial Humanities Benchmark (AHB) — Table 5sex_related_crimes#22 / 31↑41.7Source ↗official
Adversarial Humanities Benchmark (AHB) — Table 5sexual_content#27 / 31↑22.7Source ↗official
Adversarial Humanities Benchmark (AHB) — Table 5specialized_advice#24 / 31↑42.38Source ↗official
Adversarial Humanities Benchmark (AHB) — Table 5suicide_self_harm#24 / 31↑27.7Source ↗official
Adversarial Humanities Benchmark (AHB) — Table 5violent_crimes#24 / 31↑37Source ↗official
MT-JailBench CrescendoXsafety_score#15 / 21↓34.59Source ↗official
SpeechMap model completioncomplete_pct#131 / 181↑44.2Source ↗official

Values evaluations

Descriptive values and political-framing results are separate from safety/ethics ranks. Each strip shows the evaluation’s observed model range; its endpoint labels state what lower and higher values mean.

The Economist World Values Survey Cultural Map

DimensionValueDistribution
Survival ↔ Self-expression2.06
Traditional ↔ Secular0.999

Moral Trolley Arena