← Models

Model profile

Qwen3 8B

Alibabadeveloper
2025-04-29release date
#238 / 267overall rank
6eval lineages
1discovery sources

Evidence summary

Qwen3 8B has an estimated overall rank of #238; its 90% source-sensitivity interval is #159–#255. Its behavior-only rank is #235; company governance moves the combined estimate to #238. Published evidence spans 6 evals and 6 of 7 behavior components. Its strongest relative result is PHARE (bias_resistance_diagnostic, #12 of 66); its weakest is Enkrypt AI Safety Leaderboard (bias_attack_non_success_rate, #255 of 260).

Compare this model

Only models sharing at least one published sub-eval are listed.

Official and reference links

Published eval results

Rank is within that sub-eval. Black marks the observed result; the grey dot marks the value implied by the global rank. Values stay on each source’s native scale.

Eval / sub-evalRankValueDistributionBetterSource
Contextual MoralChoicehuman_agreement#6 / 220.45↑ higherSource ↗official
Enkrypt AI Safety Leaderboardbias_attack_non_success_rate#255 / 2602.33↑ higherSource ↗official
Enkrypt AI Safety Leaderboardcbrn_attack_non_success_rate#250 / 26067↑ higherSource ↗official
Enkrypt AI Safety Leaderboardharmful_attack_non_success_rate#150 / 26062.22↑ higherSource ↗official
Enkrypt AI Safety Leaderboardtoxicity_attack_non_success_rate#70 / 25898.18↑ higherSource ↗official
PandaBench JBB direct-request panelsafety_rate#14 / 460.99↑ higherSource ↗official
PHAREbias_resistance_diagnostic#12 / 660.5864↑ higherSource ↗official
PHAREhallucination_resistance_diagnostic#66 / 700.6038↑ higherSource ↗official
PHAREharm_resistance_diagnostic#56 / 700.8737↑ higherSource ↗official
PHAREjailbreak_resistance_diagnostic#33 / 670.4913↑ higherSource ↗official
PropensityBenchscore#12 / 1475.2↓ lowerSource ↗official
VETO Misfired Alignmentmisfired_alignment_rate_pct#19 / 2510.3↓ lowerSource ↗official

Values evaluations

Descriptive values and political-framing results are separate from safety/ethics ranks. Each strip shows the evaluation’s observed model range; its endpoint labels state what lower and higher values mean.

UGI Political Values

DimensionValueDistribution
Political Lean-15.8
Government48
Diplomacy62.9
Economy45.5
Society60.8