← Models

Model profile

Gemma 2B It

Googledeveloper
2024-02-21release date
#158 / 267overall rank
5eval lineages

Evidence summary

Gemma 2B It has an estimated overall rank of #158; its 90% source-sensitivity interval is #42–#202. Its behavior-only rank is #166; company governance moves the combined estimate to #158. Published evidence spans 5 evals and 6 of 7 behavior components. Its strongest relative result is Enkrypt AI Safety Leaderboard (cbrn_attack_non_success_rate, #30 of 260); its weakest is Enkrypt AI Safety Leaderboard (bias_attack_non_success_rate, #250 of 260).

Compare this model

Only models sharing at least one published sub-eval are listed.

Official and reference links

Published eval results

Rank is within that sub-eval. Black marks the observed result; the grey dot marks the value implied by the global rank. Values stay on each source’s native scale.

Eval / sub-evalRankValueDistributionBetterSource
Enkrypt AI Safety Leaderboardbias_attack_non_success_rate#250 / 2604.13↑ higherSource ↗official
Enkrypt AI Safety Leaderboardcbrn_attack_non_success_rate#30 / 26094.33↑ higherSource ↗official
Enkrypt AI Safety Leaderboardharmful_attack_non_success_rate#105 / 26077.22↑ higherSource ↗official
Enkrypt AI Safety Leaderboardtoxicity_attack_non_success_rate#248 / 25870.45↑ higherSource ↗official
Large-scale Moral Machine experiment on LLMshuman_choice_distance#21 / 391.092↓ lowerSource ↗official
S-Evalbase_en_overall#9 / 2267.5↑ higherSource ↗official
SALAD-Benchattack_enhanced_human_autonomy_integrity#6 / 3356.47↑ higherSource ↗official
SALAD-Benchattack_enhanced_information_safety_harms#6 / 3338.76↑ higherSource ↗official
SALAD-Benchattack_enhanced_malicious_use#6 / 3353.59↑ higherSource ↗official
SALAD-Benchattack_enhanced_misinformation_harms#6 / 3352.3↑ higherSource ↗official
SALAD-Benchattack_enhanced_representation_toxicity#6 / 3348.11↑ higherSource ↗official
SALAD-Benchattack_enhanced_socioeconomic_harms#6 / 3351.95↑ higherSource ↗official
SALAD-Benchbase_human_autonomy_integrity#9 / 3397.03↑ higherSource ↗official
SALAD-Benchbase_information_safety_harms#5 / 3398.38↑ higherSource ↗official
SALAD-Benchbase_malicious_use#13 / 3395.98↑ higherSource ↗official
SALAD-Benchbase_misinformation_harms#8 / 3396.41↑ higherSource ↗official
SALAD-Benchbase_representation_toxicity#5 / 3395.56↑ higherSource ↗official
SALAD-Benchbase_socioeconomic_harms#16 / 3389.78↑ higherSource ↗official
SALAD-Benchmcq_human_autonomy_integrity#23 / 3320.83↑ higherSource ↗official
SALAD-Benchmcq_information_safety_harms#25 / 3314.17↑ higherSource ↗official
SALAD-Benchmcq_malicious_use#25 / 3316.15↑ higherSource ↗official
SALAD-Benchmcq_misinformation_harms#25 / 3317.86↑ higherSource ↗official
SALAD-Benchmcq_representation_toxicity#23 / 3318.65↑ higherSource ↗official
SALAD-Benchmcq_socioeconomic_harms#25 / 3316.67↑ higherSource ↗official
SORRY-Benchavg#18 / 510.19↓ lowerSource ↗official