← Models

Model profile

Mistral Medium 3.5

Mistral AIdeveloper
2026-03-31release date
#235 / 267overall rank
3eval lineages

Evidence summary

Mistral Medium 3.5 has an estimated overall rank of #235; its 90% source-sensitivity interval is #110–#263. Its behavior-only rank is #227; company governance moves the combined estimate to #235. Published evidence spans 3 evals and 4 of 7 behavior components. Its strongest relative result is PHARE (bias_resistance_diagnostic, #11 of 66); its weakest is DystopiaBench (petrov_score, #50 of 50).

Compare this model

Only models sharing at least one published sub-eval are listed.

Official and reference links

Published eval results

Rank is within that sub-eval. Black marks the observed result; the grey dot marks the value implied by the global rank. Values stay on each source’s native scale.

Eval / sub-evalRankValueDistributionBetterSource
AA-Omnisciencehallucination_rate#172 / 3110.82↓ lowerSource ↗official
DystopiaBenchbasaglia_score#50 / 5077.7↓ lowerSource ↗official
DystopiaBenchbaudrillard_score#48 / 5079.27↓ lowerSource ↗official
DystopiaBenchhuxley_score#50 / 5086.8↓ lowerSource ↗official
DystopiaBenchlaguardia_score#50 / 5076.57↓ lowerSource ↗official
DystopiaBenchorwell_score#50 / 5082.27↓ lowerSource ↗official
DystopiaBenchpetrov_score#50 / 5090.97↓ lowerSource ↗official
PHAREbias_resistance_diagnostic#11 / 660.6098↑ higherSource ↗official
PHAREhallucination_resistance_diagnostic#54 / 700.677↑ higherSource ↗official
PHAREharm_resistance_diagnostic#46 / 700.9115↑ higherSource ↗official
PHAREjailbreak_resistance_diagnostic#55 / 670.3746↑ higherSource ↗official

Values evaluations

Descriptive values and political-framing results are separate from safety/ethics ranks. Each strip shows the evaluation’s observed model range; its endpoint labels state what lower and higher values mean.

UGI Political Values

DimensionValueDistribution
Political Lean-21.9
Government46.5
Diplomacy66.5
Economy45.6
Society60