← Models

Model profile

InternLM2 Chat 7B

InternLMdeveloper
2024-01-17release date
#101 / 333Safety rank
#417 / 645Freedom rank

Evidence summary

Safety. InternLM2 Chat 7B has an estimated Safety rank of #101; its 90% source-sensitivity interval is #26–#276. Its behavior-only rank is #100; company governance moves the combined estimate to #101. Published Safety evidence spans 4 eval lineages and 4 of 7 components. Its strongest relative result is SALAD-Bench (base_misinformation_harms, #2 of 33); its weakest is ChineseSafe (score, #17 of 22).

Freedom. InternLM2 Chat 7B has an estimated Freedom rank of #417; its 90% source-sensitivity interval is #251–#513. Published Freedom evidence spans 2 eval lineages and 1 of 1 components. Its strongest relative result is SALAD-Bench (base_information_safety_harms, #16 of 33); its weakest is SALAD-Bench (base_representation_toxicity, #31 of 33).

Compare this model

Only models sharing at least one published sub-eval are listed.

Official and reference links

Safety evals

Rank and direction are specific to the Safety portfolio. Black marks the observed result; the grey dot marks the value implied by the Safety rank. Values stay on each source’s native scale.

Eval / sub-evalRankValueDistributionSource
ChineseSafescore#17 / 2249.49Source ↗official
CMoralEvalfamilial_morality#5 / 260.51Source ↗official
CMoralEvalinternet_ethics#4 / 260.51Source ↗official
CMoralEvalpersonal_morality#4 / 260.5Source ↗official
CMoralEvalprofessional_ethics#4 / 260.52Source ↗official
CMoralEvalsocial_morality#4 / 260.51Source ↗official
JailBenchjailbreak_success_rate#5 / 1451.22Source ↗official
SALAD-Benchattack_enhanced_human_autonomy_integrity#11 / 3321.12Source ↗official
SALAD-Benchattack_enhanced_information_safety_harms#14 / 3317.59Source ↗official
SALAD-Benchattack_enhanced_malicious_use#8 / 3324.31Source ↗official
SALAD-Benchattack_enhanced_misinformation_harms#16 / 3316.45Source ↗official
SALAD-Benchattack_enhanced_representation_toxicity#13 / 3319.27Source ↗official
SALAD-Benchattack_enhanced_socioeconomic_harms#12 / 3320.35Source ↗official
SALAD-Benchbase_human_autonomy_integrity#7 / 3398.02Source ↗official
SALAD-Benchbase_information_safety_harms#18 / 3394.45Source ↗official
SALAD-Benchbase_malicious_use#3 / 3398.63Source ↗official
SALAD-Benchbase_misinformation_harms#2 / 3398.67Source ↗official
SALAD-Benchbase_representation_toxicity#3 / 3397.19Source ↗official
SALAD-Benchbase_socioeconomic_harms#5 / 3394.71Source ↗official
SALAD-Benchmcq_human_autonomy_integrity#5 / 3363.89Source ↗official
SALAD-Benchmcq_information_safety_harms#5 / 3357.5Source ↗official
SALAD-Benchmcq_malicious_use#5 / 3362.24Source ↗official
SALAD-Benchmcq_misinformation_harms#5 / 3362.62Source ↗official
SALAD-Benchmcq_representation_toxicity#5 / 3361.25Source ↗official
SALAD-Benchmcq_socioeconomic_harms#5 / 3356.11Source ↗official

Freedom evals

Rank and direction are specific to the Freedom portfolio. Black marks the observed result; the grey dot marks the value implied by the Freedom rank. Values stay on each source’s native scale.

Eval / sub-evalRankValueDistributionSource
JailBenchjailbreak_success_rate#10 / 1451.22Source ↗official
SALAD-Benchattack_enhanced_human_autonomy_integrity#22 / 3321.12Source ↗official
SALAD-Benchattack_enhanced_information_safety_harms#19 / 3317.59Source ↗official
SALAD-Benchattack_enhanced_malicious_use#26 / 3324.31Source ↗official
SALAD-Benchattack_enhanced_misinformation_harms#18 / 3316.45Source ↗official
SALAD-Benchattack_enhanced_representation_toxicity#21 / 3319.27Source ↗official
SALAD-Benchattack_enhanced_socioeconomic_harms#22 / 3320.35Source ↗official
SALAD-Benchbase_human_autonomy_integrity#27 / 3398.02Source ↗official
SALAD-Benchbase_information_safety_harms#16 / 3394.45Source ↗official
SALAD-Benchbase_malicious_use#31 / 3398.63Source ↗official
SALAD-Benchbase_misinformation_harms#31 / 3398.67Source ↗official
SALAD-Benchbase_representation_toxicity#31 / 3397.19Source ↗official
SALAD-Benchbase_socioeconomic_harms#29 / 3394.71Source ↗official