← Models

Model profile

MiMo v2 Pro

Xiaomideveloper
2026-03-18release date
#62 / 331Safety rank
#468 / 631Freedom rank

Evidence summary

Safety. MiMo v2 Pro has an estimated Safety rank of #62; its 90% source-sensitivity interval is #8–#216. Its behavior-only rank is #57; company governance moves the combined estimate to #62. Published Safety evidence spans 4 eval lineages and 5 of 7 components. Its strongest relative result is SM-Bench (anti_hallucination, #1 of 84); its weakest is WildClawBench Safety & Alignment (OpenClaw harness) (safety_alignment_score_pct, #15 of 24).

Freedom. MiMo v2 Pro has an estimated Freedom rank of #468; its 90% source-sensitivity interval is #48–#631. Published Freedom evidence spans 2 eval lineages and 1 of 1 components. Its strongest relative result is SM-Bench (eq_boundaries, #4 of 84); its weakest is SpeechMap model completion (complete_pct, #178 of 181).

Compare this model

Only models sharing at least one published sub-eval are listed.

Official and reference links

Safety evals

Rank and direction are specific to the Safety portfolio. Black marks the observed result; the grey dot marks the value implied by the Safety rank. Values stay on each source’s native scale.

Eval / sub-evalRankValueDistributionSource
AA-Omnisciencehallucination_rate#30 / 3380.2998Source ↗official
Arena Factuality — Text Arena (factuality-only weighting)factuality_bt_rating#57 / 1111435.0Source ↗official
SM-Benchadversarial#9 / 8488.29Source ↗official
SM-Benchambiguous_interpretation#46 / 8485.12Source ↗official
SM-Benchanti_hallucination#1 / 84100Source ↗official
SM-Bencheq_boundaries#4 / 8479.78Source ↗official
SM-Benchoverfit#16 / 8490.71Source ↗official
WildClawBench Safety & Alignment (OpenClaw harness)safety_alignment_score_pct#15 / 2437.48Source ↗official

Freedom evals

Rank and direction are specific to the Freedom portfolio. Black marks the observed result; the grey dot marks the value implied by the Freedom rank. Values stay on each source’s native scale.

Eval / sub-evalRankValueDistributionSource
SM-Benchadversarial#74 / 8488.29Source ↗official
SM-Bencheq_boundaries#4 / 8479.78Source ↗official
SM-Benchoverfit#16 / 8490.71Source ↗official
SpeechMap model completioncomplete_pct#178 / 18125.4Source ↗official

Values evaluations

Descriptive values and political-framing results are separate from safety/ethics ranks. Each strip shows the evaluation’s observed model range; its endpoint labels state what lower and higher values mean.

CCPBench political narrative alignment

DimensionValueDistribution
CCP-narrative alignment — all questions3.13
CCP-narrative alignment — China topics3.62
CCP-narrative alignment — non-China controls1.66