← Models

Model profile

Qwen3.6 27B

2026-04-21release date
2eval lineages

Evidence summary

Published evidence spans 2 evals and 2 of 7 behavior components. Its strongest relative result is AA-Omniscience (hallucination_rate, #45 of 311); its weakest is ODCV-Bench (misalignment_rate, #10 of 12).

Compare this model

Only models sharing at least one published sub-eval are listed.

Official and reference links

Published eval results

Rank is within that sub-eval. Black marks the observed result; the grey dot marks the value implied by the global rank. Values stay on each source’s native scale.

Eval / sub-evalRankValueDistributionBetterSource
AA-Omnisciencehallucination_rate#45 / 3110.4828↓ lowerSource ↗official
ODCV-Benchaverage_severity#9 / 121.669↓ lowerSource ↗official
ODCV-Benchmisalignment_rate#10 / 1243.75↓ lowerSource ↗official

Values evaluations

Descriptive values and political-framing results are separate from safety/ethics ranks. Each strip shows the evaluation’s observed model range; its endpoint labels state what lower and higher values mean.

UGI Political Values

DimensionValueDistribution
Political Lean-22.2
Government46.6
Diplomacy67.5
Economy43.4
Society62.9