← Models

Model profile

Minimax M2.1

MiniMaxdeveloper
2025-12-20release date
#236 / 309overall rank
6eval lineages

Evidence summary

Minimax M2.1 has an estimated overall rank of #236; its 90% source-sensitivity interval is #77–#278. Its behavior-only rank is #238; company governance moves the combined estimate to #236. Published evidence spans 6 evals and 3 of 7 behavior components. Its strongest relative result is SM-Bench (eq_boundaries, #23 of 79); its weakest is SM-Bench (anti_hallucination, #69 of 79).

Compare this model

Only models sharing at least one published sub-eval are listed.

Official and reference links

Published eval results

Rank is within that sub-eval. Black marks the observed result; the grey dot marks the value implied by the global rank. Values stay on each source’s native scale.

Eval / sub-evalRankValueDistributionSource
AA-Omnisciencehallucination_rate#103 / 3280.6854Source ↗official
Arena Factuality — Text Arena (factuality-only weighting)factuality_bt_rating#41 / 1121444.0Source ↗official
Cisco AI Defense Rolling Single-Turn Leaderboardsingle_turn_attack_success_rate#34 / 10418.59Source ↗official
HUMAINE Trust, Ethics and Safetytrust_ethics_safety_score#39 / 5425.69Source ↗official
SM-Benchadversarial#53 / 7979.76Source ↗official
SM-Benchambiguous_interpretation#61 / 7979.17Source ↗official
SM-Benchanti_hallucination#69 / 7979.06Source ↗official
SM-Bencheq_boundaries#23 / 7968.82Source ↗official
SM-Benchoverfit#46 / 7966.67Source ↗official
Vectara HHEM Factual Consistencyfactual_consistency_rate#69 / 9488.2Source ↗official

Values evaluations

Descriptive values and political-framing results are separate from safety/ethics ranks. Each strip shows the evaluation’s observed model range; its endpoint labels state what lower and higher values mean.

UGI Political Values

DimensionValueDistribution
Political Lean-20.3
Government49.7
Diplomacy63.9
Economy46.7
Society59.4