← Models

Model profile

Mistral Medium 3.5

Mistral AIdeveloper
2026-03-31release date
#252 / 309overall rank
4eval lineages

Evidence summary

Mistral Medium 3.5 has an estimated overall rank of #252; its 90% source-sensitivity interval is #120–#297. Its behavior-only rank is #243; company governance moves the combined estimate to #252. Published evidence spans 4 evals and 4 of 7 behavior components. Its strongest relative result is PHARE (bias_resistance_diagnostic, #11 of 66); its weakest is DystopiaBench (petrov_score, #50 of 50).

Compare this model

Only models sharing at least one published sub-eval are listed.

Official and reference links

Published eval results

Rank is within that sub-eval. Black marks the observed result; the grey dot marks the value implied by the global rank. Values stay on each source’s native scale.

Eval / sub-evalRankValueDistributionSource
AA-Omnisciencehallucination_rate#163 / 3280.8159Source ↗official
Arena Factuality — Text Arena (factuality-only weighting)factuality_bt_rating#62 / 1121433.0Source ↗official
DystopiaBenchbasaglia_score#50 / 5077.7Source ↗official
DystopiaBenchbaudrillard_score#48 / 5079.27Source ↗official
DystopiaBenchhuxley_score#50 / 5086.8Source ↗official
DystopiaBenchlaguardia_score#50 / 5076.57Source ↗official
DystopiaBenchorwell_score#50 / 5082.27Source ↗official
DystopiaBenchpetrov_score#50 / 5090.97Source ↗official
PHAREbias_resistance_diagnostic#11 / 660.6098Source ↗official
PHAREhallucination_resistance_diagnostic#54 / 700.677Source ↗official
PHAREharm_resistance_diagnostic#46 / 700.9115Source ↗official
PHAREjailbreak_resistance_diagnostic#55 / 670.3746Source ↗official

Values evaluations

Descriptive values and political-framing results are separate from safety/ethics ranks. Each strip shows the evaluation’s observed model range; its endpoint labels state what lower and higher values mean.

UGI Political Values

DimensionValueDistribution
Political Lean-21.9
Government46.5
Diplomacy66.5
Economy45.6
Society60