← Models

Model profile

InternLM2 Chat 20B

InternLMdeveloper
2024-01-17release date
#123 / 333Safety rank
#459 / 645Freedom rank

Evidence summary

Safety. InternLM2 Chat 20B has an estimated Safety rank of #123; its 90% source-sensitivity interval is #36–#266. Its behavior-only rank is #121; company governance moves the combined estimate to #123. Published Safety evidence spans 5 eval lineages and 5 of 7 components. Its strongest relative result is SALAD-Bench (base_representation_toxicity, #2 of 33); its weakest is ChineseSafe (score, #15 of 22).

Freedom. InternLM2 Chat 20B has an estimated Freedom rank of #459; its 90% source-sensitivity interval is #246–#545. Published Freedom evidence spans 2 eval lineages and 1 of 1 components. Its strongest relative result is SALAD-Bench (attack_enhanced_information_safety_harms, #12 of 33); its weakest is SALAD-Bench (base_representation_toxicity, #32 of 33).

Compare this model

Only models sharing at least one published sub-eval are listed.

Official and reference links

Safety evals

Rank and direction are specific to the Safety portfolio. Black marks the observed result; the grey dot marks the value implied by the Safety rank. Values stay on each source’s native scale.

Eval / sub-evalRankValueDistributionSource
ChineseSafescore#15 / 2253.67Source ↗official
CMoralEvalfamilial_morality#3 / 260.56Source ↗official
CMoralEvalinternet_ethics#3 / 260.54Source ↗official
CMoralEvalpersonal_morality#3 / 260.52Source ↗official
CMoralEvalprofessional_ethics#3 / 260.54Source ↗official
CMoralEvalsocial_morality#3 / 260.54Source ↗official
Enkrypt AI Safety Leaderboardbias_attack_non_success_rate#86 / 24820.41Source ↗official
Enkrypt AI Safety Leaderboardcbrn_attack_non_success_rate#35 / 24893.33Source ↗official
Enkrypt AI Safety Leaderboardharmful_attack_non_success_rate#123 / 24871.67Source ↗official
Enkrypt AI Safety Leaderboardtoxicity_attack_non_success_rate#121 / 24696.41Source ↗official
FinEval Financial Security Knowledgefinancial_security_accuracy_pct#10 / 1973.1Source ↗official
SALAD-Benchattack_enhanced_human_autonomy_integrity#15 / 3315.52Source ↗official
SALAD-Benchattack_enhanced_information_safety_harms#21 / 337.17Source ↗official
SALAD-Benchattack_enhanced_malicious_use#15 / 3312.56Source ↗official
SALAD-Benchattack_enhanced_misinformation_harms#19 / 339.38Source ↗official
SALAD-Benchattack_enhanced_representation_toxicity#20 / 3310.12Source ↗official
SALAD-Benchattack_enhanced_socioeconomic_harms#17 / 3312.55Source ↗official
SALAD-Benchbase_human_autonomy_integrity#3 / 3398.54Source ↗official
SALAD-Benchbase_information_safety_harms#11 / 3396.68Source ↗official
SALAD-Benchbase_malicious_use#2 / 3398.88Source ↗official
SALAD-Benchbase_misinformation_harms#2 / 3398.67Source ↗official
SALAD-Benchbase_representation_toxicity#2 / 3397.53Source ↗official
SALAD-Benchbase_socioeconomic_harms#3 / 3395.77Source ↗official
SALAD-Benchmcq_human_autonomy_integrity#8 / 3361.11Source ↗official
SALAD-Benchmcq_information_safety_harms#7 / 3351.11Source ↗official
SALAD-Benchmcq_malicious_use#6 / 3357.24Source ↗official
SALAD-Benchmcq_misinformation_harms#6 / 3359.05Source ↗official
SALAD-Benchmcq_representation_toxicity#6 / 3357.92Source ↗official
SALAD-Benchmcq_socioeconomic_harms#10 / 3346.11Source ↗official

Freedom evals

Rank and direction are specific to the Freedom portfolio. Black marks the observed result; the grey dot marks the value implied by the Freedom rank. Values stay on each source’s native scale.

Eval / sub-evalRankValueDistributionSource
Enkrypt AI Safety Leaderboardcbrn_attack_non_success_rate#212 / 24893.33Source ↗official
Enkrypt AI Safety Leaderboardharmful_attack_non_success_rate#124 / 24871.67Source ↗official
Enkrypt AI Safety Leaderboardtoxicity_attack_non_success_rate#126 / 24696.41Source ↗official
SALAD-Benchattack_enhanced_human_autonomy_integrity#19 / 3315.52Source ↗official
SALAD-Benchattack_enhanced_information_safety_harms#12 / 337.17Source ↗official
SALAD-Benchattack_enhanced_malicious_use#19 / 3312.56Source ↗official
SALAD-Benchattack_enhanced_misinformation_harms#14 / 339.38Source ↗official
SALAD-Benchattack_enhanced_representation_toxicity#14 / 3310.12Source ↗official
SALAD-Benchattack_enhanced_socioeconomic_harms#17 / 3312.55Source ↗official
SALAD-Benchbase_human_autonomy_integrity#31 / 3398.54Source ↗official
SALAD-Benchbase_information_safety_harms#23 / 3396.68Source ↗official
SALAD-Benchbase_malicious_use#32 / 3398.88Source ↗official
SALAD-Benchbase_misinformation_harms#31 / 3398.67Source ↗official
SALAD-Benchbase_representation_toxicity#32 / 3397.53Source ↗official
SALAD-Benchbase_socioeconomic_harms#31 / 3395.77Source ↗official