← Models

Model profile

InternLM Chat 7B

InternLMdeveloper
2023-07-06release date
#209 / 333Safety rank
#308 / 645Freedom rank

Evidence summary

Safety. InternLM Chat 7B has an estimated Safety rank of #209; its 90% source-sensitivity interval is #63–#279. Its behavior-only rank is #213; company governance moves the combined estimate to #209. Published Safety evidence spans 5 eval lineages and 4 of 7 components. Its strongest relative result is FLAMES (legality, #1 of 13); its weakest is SALAD-Bench (mcq_representation_toxicity, #30 of 33).

Freedom. InternLM Chat 7B has an estimated Freedom rank of #308; its 90% source-sensitivity interval is #76–#537. Published Freedom evidence spans 4 eval lineages and 1 of 1 components. Its strongest relative result is SuperCLUE Safety (traditional_safety, #2 of 31); its weakest is FLAMES (legality, #13 of 13).

Compare this model

Only models sharing at least one published sub-eval are listed.

Official and reference links

Safety evals

Rank and direction are specific to the Safety portfolio. Black marks the observed result; the grey dot marks the value implied by the Safety rank. Values stay on each source’s native scale.

Eval / sub-evalRankValueDistributionSource
Fake Alignment (FINE)multiple_choice_safe_decision_rate#6 / 1457.33Source ↗official
Fake Alignment (FINE)open_ended_safe_response_rate#11 / 1492Source ↗official
FLAMESdata_protection#2 / 1361.84Source ↗official
FLAMESfairness#2 / 1344.58Source ↗official
FLAMESlegality#1 / 1376.09Source ↗official
FLAMESmorality#3 / 1351.24Source ↗official
FLAMESsafety#5 / 1335.9Source ↗official
SafetyBenchEM#6 / 2175.4Source ↗official
SafetyBenchIA#8 / 2179.5Source ↗official
SafetyBenchMH#6 / 2184.3Source ↗official
SafetyBenchOFF#12 / 2167.2Source ↗official
SafetyBenchPH#7 / 2174.15Source ↗official
SafetyBenchPP#7 / 2178.7Source ↗official
SafetyBenchUB#8 / 2164.75Source ↗official
SALAD-Benchattack_enhanced_human_autonomy_integrity#11 / 3321.12Source ↗official
SALAD-Benchattack_enhanced_information_safety_harms#16 / 3316.61Source ↗official
SALAD-Benchattack_enhanced_malicious_use#12 / 3322.51Source ↗official
SALAD-Benchattack_enhanced_misinformation_harms#11 / 3319.9Source ↗official
SALAD-Benchattack_enhanced_representation_toxicity#9 / 3323.71Source ↗official
SALAD-Benchattack_enhanced_socioeconomic_harms#11 / 3322.08Source ↗official
SALAD-Benchbase_human_autonomy_integrity#10 / 3396.85Source ↗official
SALAD-Benchbase_information_safety_harms#15 / 3395.13Source ↗official
SALAD-Benchbase_malicious_use#12 / 3396.28Source ↗official
SALAD-Benchbase_misinformation_harms#5 / 3397.05Source ↗official
SALAD-Benchbase_representation_toxicity#7 / 3394.37Source ↗official
SALAD-Benchbase_socioeconomic_harms#15 / 3390.95Source ↗official
SALAD-Benchmcq_human_autonomy_integrity#30 / 330Source ↗official
SALAD-Benchmcq_information_safety_harms#30 / 330Source ↗official
SALAD-Benchmcq_malicious_use#30 / 330.0641Source ↗official
SALAD-Benchmcq_misinformation_harms#30 / 330Source ↗official
SALAD-Benchmcq_representation_toxicity#30 / 330.1042Source ↗official
SALAD-Benchmcq_socioeconomic_harms#30 / 330Source ↗official
SuperCLUE Safetyinstruction_attack#20 / 3158.62Source ↗official
SuperCLUE Safetyresponsible_ai#28 / 3145.45Source ↗official
SuperCLUE Safetytraditional_safety#28 / 3165.96Source ↗official

Freedom evals

Rank and direction are specific to the Freedom portfolio. Black marks the observed result; the grey dot marks the value implied by the Freedom rank. Values stay on each source’s native scale.

Eval / sub-evalRankValueDistributionSource
Fake Alignment (FINE)open_ended_safe_response_rate#4 / 1492Source ↗official
FLAMESdata_protection#12 / 1361.84Source ↗official
FLAMESfairness#12 / 1344.58Source ↗official
FLAMESlegality#13 / 1376.09Source ↗official
FLAMESmorality#11 / 1351.24Source ↗official
FLAMESsafety#9 / 1335.9Source ↗official
SALAD-Benchattack_enhanced_human_autonomy_integrity#22 / 3321.12Source ↗official
SALAD-Benchattack_enhanced_information_safety_harms#18 / 3316.61Source ↗official
SALAD-Benchattack_enhanced_malicious_use#22 / 3322.51Source ↗official
SALAD-Benchattack_enhanced_misinformation_harms#23 / 3319.9Source ↗official
SALAD-Benchattack_enhanced_representation_toxicity#25 / 3323.71Source ↗official
SALAD-Benchattack_enhanced_socioeconomic_harms#23 / 3322.08Source ↗official
SALAD-Benchbase_human_autonomy_integrity#24 / 3396.85Source ↗official
SALAD-Benchbase_information_safety_harms#19 / 3395.13Source ↗official
SALAD-Benchbase_malicious_use#22 / 3396.28Source ↗official
SALAD-Benchbase_misinformation_harms#29 / 3397.05Source ↗official
SALAD-Benchbase_representation_toxicity#27 / 3394.37Source ↗official
SALAD-Benchbase_socioeconomic_harms#19 / 3390.95Source ↗official
SuperCLUE Safetyinstruction_attack#9 / 3158.62Source ↗official
SuperCLUE Safetytraditional_safety#2 / 3165.96Source ↗official