← Models

Model profile

InternLM Chat 20B

InternLMdeveloper
2023-09-18release date
#112 / 333Safety rank
#448 / 645Freedom rank

Evidence summary

Safety. InternLM Chat 20B has an estimated Safety rank of #112; its 90% source-sensitivity interval is #36–#235. Its behavior-only rank is #111; company governance moves the combined estimate to #112. Published Safety evidence spans 4 eval lineages and 4 of 7 components. Its strongest relative result is SALAD-Bench (base_information_safety_harms, #1 of 33); its weakest is SALAD-Bench (mcq_representation_toxicity, #29 of 33).

Freedom. InternLM Chat 20B has an estimated Freedom rank of #448; its 90% source-sensitivity interval is #176–#615. Published Freedom evidence spans 4 eval lineages and 1 of 1 components. Its strongest relative result is SuperCLUE Safety (traditional_safety, #5 of 31); its weakest is FLAMES (fairness, #13 of 13).

Compare this model

Only models sharing at least one published sub-eval are listed.

Official and reference links

Safety evals

Rank and direction are specific to the Safety portfolio. Black marks the observed result; the grey dot marks the value implied by the Safety rank. Values stay on each source’s native scale.

Eval / sub-evalRankValueDistributionSource
Fake Alignment (FINE)multiple_choice_safe_decision_rate#3 / 1469.33Source ↗official
Fake Alignment (FINE)open_ended_safe_response_rate#7 / 1496Source ↗official
FLAMESdata_protection#1 / 1363.16Source ↗official
FLAMESfairness#1 / 1352.61Source ↗official
FLAMESlegality#2 / 1371.74Source ↗official
FLAMESmorality#1 / 1354.23Source ↗official
FLAMESsafety#3 / 1351.05Source ↗official
SALAD-Benchattack_enhanced_human_autonomy_integrity#7 / 3329.74Source ↗official
SALAD-Benchattack_enhanced_information_safety_harms#9 / 3323.45Source ↗official
SALAD-Benchattack_enhanced_malicious_use#7 / 3329.36Source ↗official
SALAD-Benchattack_enhanced_misinformation_harms#8 / 3327.96Source ↗official
SALAD-Benchattack_enhanced_representation_toxicity#7 / 3331.79Source ↗official
SALAD-Benchattack_enhanced_socioeconomic_harms#7 / 3327.27Source ↗official
SALAD-Benchbase_human_autonomy_integrity#2 / 3398.89Source ↗official
SALAD-Benchbase_information_safety_harms#1 / 3399.8Source ↗official
SALAD-Benchbase_malicious_use#5 / 3398.44Source ↗official
SALAD-Benchbase_misinformation_harms#6 / 3397Source ↗official
SALAD-Benchbase_representation_toxicity#8 / 3393.39Source ↗official
SALAD-Benchbase_socioeconomic_harms#2 / 3396.36Source ↗official
SALAD-Benchmcq_human_autonomy_integrity#29 / 333.333Source ↗official
SALAD-Benchmcq_information_safety_harms#27 / 334.167Source ↗official
SALAD-Benchmcq_malicious_use#29 / 333.974Source ↗official
SALAD-Benchmcq_misinformation_harms#29 / 333.333Source ↗official
SALAD-Benchmcq_representation_toxicity#29 / 333.229Source ↗official
SALAD-Benchmcq_socioeconomic_harms#27 / 337.778Source ↗official
SuperCLUE Safetyinstruction_attack#9 / 3168.97Source ↗official
SuperCLUE Safetyresponsible_ai#11 / 3165.45Source ↗official
SuperCLUE Safetytraditional_safety#25 / 3168.09Source ↗official

Freedom evals

Rank and direction are specific to the Freedom portfolio. Black marks the observed result; the grey dot marks the value implied by the Freedom rank. Values stay on each source’s native scale.

Eval / sub-evalRankValueDistributionSource
Fake Alignment (FINE)open_ended_safe_response_rate#7 / 1496Source ↗official
FLAMESdata_protection#13 / 1363.16Source ↗official
FLAMESfairness#13 / 1352.61Source ↗official
FLAMESlegality#12 / 1371.74Source ↗official
FLAMESmorality#12 / 1354.23Source ↗official
FLAMESsafety#11 / 1351.05Source ↗official
SALAD-Benchattack_enhanced_human_autonomy_integrity#27 / 3329.74Source ↗official
SALAD-Benchattack_enhanced_information_safety_harms#25 / 3323.45Source ↗official
SALAD-Benchattack_enhanced_malicious_use#27 / 3329.36Source ↗official
SALAD-Benchattack_enhanced_misinformation_harms#26 / 3327.96Source ↗official
SALAD-Benchattack_enhanced_representation_toxicity#27 / 3331.79Source ↗official
SALAD-Benchattack_enhanced_socioeconomic_harms#27 / 3327.27Source ↗official
SALAD-Benchbase_human_autonomy_integrity#32 / 3398.89Source ↗official
SALAD-Benchbase_information_safety_harms#33 / 3399.8Source ↗official
SALAD-Benchbase_malicious_use#29 / 3398.44Source ↗official
SALAD-Benchbase_misinformation_harms#28 / 3397Source ↗official
SALAD-Benchbase_representation_toxicity#26 / 3393.39Source ↗official
SALAD-Benchbase_socioeconomic_harms#32 / 3396.36Source ↗official
SuperCLUE Safetyinstruction_attack#18 / 3168.97Source ↗official
SuperCLUE Safetytraditional_safety#5 / 3168.09Source ↗official