← Models

Model profile

Gemma 7B It

Googledeveloper
2024-02-21release date
#185 / 305overall rank
7eval lineages

Evidence summary

Gemma 7B It has an estimated overall rank of #185; its 90% source-sensitivity interval is #58–#237. Its behavior-only rank is #192; company governance moves the combined estimate to #185. Published evidence spans 7 evals and 7 of 7 behavior components. Its strongest relative result is SALAD-Bench (base_representation_toxicity, #6 of 33); its weakest is Enkrypt AI Safety Leaderboard (toxicity_attack_non_success_rate, #224 of 239).

Compare this model

Only models sharing at least one published sub-eval are listed.

Official and reference links

Published eval results

Rank is within that sub-eval. Black marks the observed result; the grey dot marks the value implied by the global rank. Values stay on each source’s native scale.

Eval / sub-evalRankValueDistributionSource
Enkrypt AI Safety Leaderboardbias_attack_non_success_rate#82 / 24120.41Source ↗official
Enkrypt AI Safety Leaderboardcbrn_attack_non_success_rate#206 / 24178.83Source ↗official
Enkrypt AI Safety Leaderboardharmful_attack_non_success_rate#164 / 24154.44Source ↗official
Enkrypt AI Safety Leaderboardtoxicity_attack_non_success_rate#224 / 23977.05Source ↗official
Large-scale Moral Machine experiment on LLMshuman_choice_distance#32 / 391.314Source ↗official
Microsoft Phi Safety Panelsharmful_continuation#6 / 100.013Source ↗official
Microsoft Phi Safety Panelsharmful_summarization#3 / 100.103Source ↗official
Microsoft Phi Safety Panelsjailbreak#4 / 100.114Source ↗official
Microsoft Phi Safety Panelsthird_party_harm#8 / 100.383Source ↗official
OR-Benchover_refusal_rate#7 / 2526.3Source ↗official
OR-Benchtoxic_acceptance_rate#17 / 2514.5Source ↗official
S-Evalbase_en_overall#10 / 2261.8Source ↗official
SALAD-Benchattack_enhanced_human_autonomy_integrity#16 / 3313.36Source ↗official
SALAD-Benchattack_enhanced_information_safety_harms#11 / 3322.8Source ↗official
SALAD-Benchattack_enhanced_malicious_use#19 / 339.95Source ↗official
SALAD-Benchattack_enhanced_misinformation_harms#12 / 3318.91Source ↗official
SALAD-Benchattack_enhanced_representation_toxicity#15 / 3317.56Source ↗official
SALAD-Benchattack_enhanced_socioeconomic_harms#18 / 3312.12Source ↗official
SALAD-Benchbase_human_autonomy_integrity#17 / 3394.82Source ↗official
SALAD-Benchbase_information_safety_harms#9 / 3397.49Source ↗official
SALAD-Benchbase_malicious_use#17 / 3393.54Source ↗official
SALAD-Benchbase_misinformation_harms#12 / 3395.57Source ↗official
SALAD-Benchbase_representation_toxicity#6 / 3394.42Source ↗official
SALAD-Benchbase_socioeconomic_harms#22 / 3386.13Source ↗official
SALAD-Benchmcq_human_autonomy_integrity#15 / 3340.56Source ↗official
SALAD-Benchmcq_information_safety_harms#12 / 3340.56Source ↗official
SALAD-Benchmcq_malicious_use#14 / 3340.38Source ↗official
SALAD-Benchmcq_misinformation_harms#16 / 3338.1Source ↗official
SALAD-Benchmcq_representation_toxicity#16 / 3338.85Source ↗official
SALAD-Benchmcq_socioeconomic_harms#14 / 3338.89Source ↗official
SORRY-Benchavg#16 / 510.18Source ↗official