← Models

Model profile

Mixtral 8X7B Instruct

Mistral AIdeveloper
2023-12-10release date
#260 / 267overall rank
10eval lineages

Evidence summary

Mixtral 8X7B Instruct has an estimated overall rank of #260; its 90% source-sensitivity interval is #188–#265. Its behavior-only rank is #254; company governance moves the combined estimate to #260. Published evidence spans 10 evals and 6 of 7 behavior components. Its strongest relative result is SALAD-Bench (mcq_information_safety_harms, #6 of 33); its weakest is HELM Safety (harmbench, #77 of 80).

Compare this model

Only models sharing at least one published sub-eval are listed.

Official and reference links

Published eval results

Rank is within that sub-eval. Black marks the observed result; the grey dot marks the value implied by the global rank. Values stay on each source’s native scale.

Eval / sub-evalRankValueDistributionBetterSource
AIRBench 2024 Safety Scenariossafety_scenarios#74 / 800.391↑ higherSource ↗official
CASE-Benchagreement_accuracy#4 / 783.11↑ higherSource ↗official
Contextual MoralChoicehuman_agreement#6 / 220.45↑ higherSource ↗official
Enkrypt AI Safety Leaderboardbias_attack_non_success_rate#209 / 26011.37↑ higherSource ↗official
Enkrypt AI Safety Leaderboardcbrn_attack_non_success_rate#162 / 26086.67↑ higherSource ↗official
Enkrypt AI Safety Leaderboardharmful_attack_non_success_rate#229 / 26038.33↑ higherSource ↗official
Enkrypt AI Safety Leaderboardtoxicity_attack_non_success_rate#164 / 25895.18↑ higherSource ↗official
HarmBenchdr#25 / 2847.3↓ lowerSource ↗official
HELM Safetyanthropic_red_team#75 / 800.928↑ higherSource ↗official
HELM Safetybbq#63 / 800.857↑ higherSource ↗official
HELM Safetyharmbench#77 / 800.451↑ higherSource ↗official
HELM Safetysimple_safety_tests#75 / 800.905↑ higherSource ↗official
HELM Safetyxstest#61 / 800.931↑ higherSource ↗official
SALAD-Benchattack_enhanced_human_autonomy_integrity#19 / 3311.85↑ higherSource ↗official
SALAD-Benchattack_enhanced_information_safety_harms#25 / 335.21↑ higherSource ↗official
SALAD-Benchattack_enhanced_malicious_use#25 / 337.67↑ higherSource ↗official
SALAD-Benchattack_enhanced_misinformation_harms#18 / 339.54↑ higherSource ↗official
SALAD-Benchattack_enhanced_representation_toxicity#19 / 3310.35↑ higherSource ↗official
SALAD-Benchattack_enhanced_socioeconomic_harms#21 / 339.52↑ higherSource ↗official
SALAD-Benchbase_human_autonomy_integrity#28 / 3376↑ higherSource ↗official
SALAD-Benchbase_information_safety_harms#27 / 3386.8↑ higherSource ↗official
SALAD-Benchbase_malicious_use#28 / 3367.65↑ higherSource ↗official
SALAD-Benchbase_misinformation_harms#28 / 3384.39↑ higherSource ↗official
SALAD-Benchbase_representation_toxicity#25 / 3382.05↑ higherSource ↗official
SALAD-Benchbase_socioeconomic_harms#27 / 3380.85↑ higherSource ↗official
SALAD-Benchmcq_human_autonomy_integrity#9 / 3351.39↑ higherSource ↗official
SALAD-Benchmcq_information_safety_harms#6 / 3351.94↑ higherSource ↗official
SALAD-Benchmcq_malicious_use#9 / 3353.27↑ higherSource ↗official
SALAD-Benchmcq_misinformation_harms#9 / 3352.86↑ higherSource ↗official
SALAD-Benchmcq_representation_toxicity#9 / 3352.08↑ higherSource ↗official
SALAD-Benchmcq_socioeconomic_harms#8 / 3348.89↑ higherSource ↗official
SORRY-Benchavg#43 / 510.56↓ lowerSource ↗official

Values evaluations

Descriptive values and political-framing results are separate from safety/ethics ranks. Each strip shows the evaluation’s observed model range; its endpoint labels state what lower and higher values mean.

UGI Political Values

DimensionValueDistribution
Political Lean-5.8
Government45.9
Diplomacy57.2
Economy40.6
Society57.8