← Models

Model profile

Llama 2 13B Chat

Metadeveloper
2023-07-18release date
#164 / 267overall rank
8eval lineages

Evidence summary

Llama 2 13B Chat has an estimated overall rank of #164; its 90% source-sensitivity interval is #61–#214. Its behavior-only rank is #149; company governance moves the combined estimate to #164. Published evidence spans 8 evals and 6 of 7 behavior components. Its strongest relative result is OR-Bench (toxic_acceptance_rate, #2 of 25); its weakest is SafetyBench (OFF, #20 of 21).

Compare this model

Only models sharing at least one published sub-eval are listed.

Official and reference links

Published eval results

Rank is within that sub-eval. Black marks the observed result; the grey dot marks the value implied by the global rank. Values stay on each source’s native scale.

Eval / sub-evalRankValueDistributionBetterSource
Enkrypt AI Safety Leaderboardbias_attack_non_success_rate#211 / 26011.11↑ higherSource ↗official
Enkrypt AI Safety Leaderboardcbrn_attack_non_success_rate#58 / 26091.83↑ higherSource ↗official
Enkrypt AI Safety Leaderboardharmful_attack_non_success_rate#61 / 26088.33↑ higherSource ↗official
Enkrypt AI Safety Leaderboardtoxicity_attack_non_success_rate#26 / 25899.55↑ higherSource ↗official
HarmBenchdr#4 / 282.8↓ lowerSource ↗official
JailBenchjailbreak_success_rate#6 / 1455.39↓ lowerSource ↗official
OR-Benchover_refusal_rate#20 / 2591↓ lowerSource ↗official
OR-Benchtoxic_acceptance_rate#2 / 250.3↓ lowerSource ↗official
S-Evalbase_en_overall#2 / 2285.1↑ higherSource ↗official
SafetyBenchEM#17 / 2154.6↑ higherSource ↗official
SafetyBenchIA#15 / 2168.5↑ higherSource ↗official
SafetyBenchMH#15 / 2173.6↑ higherSource ↗official
SafetyBenchOFF#20 / 2148.4↑ higherSource ↗official
SafetyBenchPH#16 / 2160.7↑ higherSource ↗official
SafetyBenchPP#16 / 2170.1↑ higherSource ↗official
SafetyBenchUB#5 / 2166.3↑ higherSource ↗official
SALAD-Benchattack_enhanced_human_autonomy_integrity#4 / 3362.72↑ higherSource ↗official
SALAD-Benchattack_enhanced_information_safety_harms#4 / 3371.01↑ higherSource ↗official
SALAD-Benchattack_enhanced_malicious_use#5 / 3362.56↑ higherSource ↗official
SALAD-Benchattack_enhanced_misinformation_harms#4 / 3370.23↑ higherSource ↗official
SALAD-Benchattack_enhanced_representation_toxicity#5 / 3365.85↑ higherSource ↗official
SALAD-Benchattack_enhanced_socioeconomic_harms#4 / 3368.4↑ higherSource ↗official
SALAD-Benchbase_human_autonomy_integrity#4 / 3398.43↑ higherSource ↗official
SALAD-Benchbase_information_safety_harms#16 / 3394.92↑ higherSource ↗official
SALAD-Benchbase_malicious_use#7 / 3397.98↑ higherSource ↗official
SALAD-Benchbase_misinformation_harms#4 / 3397.64↑ higherSource ↗official
SALAD-Benchbase_representation_toxicity#4 / 3395.71↑ higherSource ↗official
SALAD-Benchbase_socioeconomic_harms#13 / 3391.19↑ higherSource ↗official
SALAD-Benchmcq_human_autonomy_integrity#26 / 3310.28↑ higherSource ↗official
SALAD-Benchmcq_information_safety_harms#23 / 3317.5↑ higherSource ↗official
SALAD-Benchmcq_malicious_use#26 / 337.885↑ higherSource ↗official
SALAD-Benchmcq_misinformation_harms#26 / 3310.95↑ higherSource ↗official
SALAD-Benchmcq_representation_toxicity#26 / 338.229↑ higherSource ↗official
SALAD-Benchmcq_socioeconomic_harms#26 / 3312.78↑ higherSource ↗official
SORRY-Benchavg#15 / 510.15↓ lowerSource ↗official