← Models

Model profile

Gemma 2B It

Googledeveloper
2024-02-21release date
#188 / 312overall rank
5eval lineages

Evidence summary

Gemma 2B It has an estimated overall rank of #188; its 90% source-sensitivity interval is #46–#250. Its behavior-only rank is #200; company governance moves the combined estimate to #188. Published evidence spans 5 evals and 6 of 7 behavior components. Its strongest relative result is Enkrypt AI Safety Leaderboard (cbrn_attack_non_success_rate, #26 of 241); its weakest is Enkrypt AI Safety Leaderboard (toxicity_attack_non_success_rate, #230 of 239).

Compare this model

Only models sharing at least one published sub-eval are listed.

Official and reference links

Published eval results

Rank is within that sub-eval. Black marks the observed result; the grey dot marks the value implied by the global rank. Values stay on each source’s native scale.

Eval / sub-evalRankValueDistributionSource
Enkrypt AI Safety Leaderboardbias_attack_non_success_rate#231 / 2414.13Source ↗official
Enkrypt AI Safety Leaderboardcbrn_attack_non_success_rate#26 / 24194.33Source ↗official
Enkrypt AI Safety Leaderboardharmful_attack_non_success_rate#93 / 24177.22Source ↗official
Enkrypt AI Safety Leaderboardtoxicity_attack_non_success_rate#230 / 23970.45Source ↗official
Large-scale Moral Machine experiment on LLMshuman_choice_distance#21 / 391.092Source ↗official
S-Evalbase_en_overall#9 / 2267.5Source ↗official
SALAD-Benchattack_enhanced_human_autonomy_integrity#6 / 3356.47Source ↗official
SALAD-Benchattack_enhanced_information_safety_harms#6 / 3338.76Source ↗official
SALAD-Benchattack_enhanced_malicious_use#6 / 3353.59Source ↗official
SALAD-Benchattack_enhanced_misinformation_harms#6 / 3352.3Source ↗official
SALAD-Benchattack_enhanced_representation_toxicity#6 / 3348.11Source ↗official
SALAD-Benchattack_enhanced_socioeconomic_harms#6 / 3351.95Source ↗official
SALAD-Benchbase_human_autonomy_integrity#9 / 3397.03Source ↗official
SALAD-Benchbase_information_safety_harms#5 / 3398.38Source ↗official
SALAD-Benchbase_malicious_use#13 / 3395.98Source ↗official
SALAD-Benchbase_misinformation_harms#8 / 3396.41Source ↗official
SALAD-Benchbase_representation_toxicity#5 / 3395.56Source ↗official
SALAD-Benchbase_socioeconomic_harms#16 / 3389.78Source ↗official
SALAD-Benchmcq_human_autonomy_integrity#23 / 3320.83Source ↗official
SALAD-Benchmcq_information_safety_harms#25 / 3314.17Source ↗official
SALAD-Benchmcq_malicious_use#25 / 3316.15Source ↗official
SALAD-Benchmcq_misinformation_harms#25 / 3317.86Source ↗official
SALAD-Benchmcq_representation_toxicity#23 / 3318.65Source ↗official
SALAD-Benchmcq_socioeconomic_harms#25 / 3316.67Source ↗official
SORRY-Benchavg#18 / 510.19Source ↗official