← Models

Model profile

Internlm Chat 7B

InternLMdeveloper
2023-07-06release date
#176 / 309overall rank
5eval lineages

Evidence summary

Internlm Chat 7B has an estimated overall rank of #176; its 90% source-sensitivity interval is #48–#255. Published evidence spans 5 evals and 4 of 7 behavior components. Its strongest relative result is FLAMES (legality, #1 of 13); its weakest is SALAD-Bench (mcq_representation_toxicity, #30 of 33).

Compare this model

Only models sharing at least one published sub-eval are listed.

Official and reference links

Published eval results

Rank is within that sub-eval. Black marks the observed result; the grey dot marks the value implied by the global rank. Values stay on each source’s native scale.

Eval / sub-evalRankValueDistributionSource
Fake Alignment (FINE)multiple_choice_safe_decision_rate#6 / 1457.33Source ↗official
Fake Alignment (FINE)open_ended_safe_response_rate#11 / 1492Source ↗official
FLAMESdata_protection#2 / 1361.84Source ↗official
FLAMESfairness#2 / 1344.58Source ↗official
FLAMESlegality#1 / 1376.09Source ↗official
FLAMESmorality#3 / 1351.24Source ↗official
FLAMESsafety#5 / 1335.9Source ↗official
SafetyBenchEM#6 / 2175.4Source ↗official
SafetyBenchIA#8 / 2179.5Source ↗official
SafetyBenchMH#6 / 2184.3Source ↗official
SafetyBenchOFF#12 / 2167.2Source ↗official
SafetyBenchPH#7 / 2174.15Source ↗official
SafetyBenchPP#7 / 2178.7Source ↗official
SafetyBenchUB#8 / 2164.75Source ↗official
SALAD-Benchattack_enhanced_human_autonomy_integrity#11 / 3321.12Source ↗official
SALAD-Benchattack_enhanced_information_safety_harms#16 / 3316.61Source ↗official
SALAD-Benchattack_enhanced_malicious_use#12 / 3322.51Source ↗official
SALAD-Benchattack_enhanced_misinformation_harms#11 / 3319.9Source ↗official
SALAD-Benchattack_enhanced_representation_toxicity#9 / 3323.71Source ↗official
SALAD-Benchattack_enhanced_socioeconomic_harms#11 / 3322.08Source ↗official
SALAD-Benchbase_human_autonomy_integrity#10 / 3396.85Source ↗official
SALAD-Benchbase_information_safety_harms#15 / 3395.13Source ↗official
SALAD-Benchbase_malicious_use#12 / 3396.28Source ↗official
SALAD-Benchbase_misinformation_harms#5 / 3397.05Source ↗official
SALAD-Benchbase_representation_toxicity#7 / 3394.37Source ↗official
SALAD-Benchbase_socioeconomic_harms#15 / 3390.95Source ↗official
SALAD-Benchmcq_human_autonomy_integrity#30 / 330Source ↗official
SALAD-Benchmcq_information_safety_harms#30 / 330Source ↗official
SALAD-Benchmcq_malicious_use#30 / 330.0641Source ↗official
SALAD-Benchmcq_misinformation_harms#30 / 330Source ↗official
SALAD-Benchmcq_representation_toxicity#30 / 330.1042Source ↗official
SALAD-Benchmcq_socioeconomic_harms#30 / 330Source ↗official
SuperCLUE Safetyinstruction_attack#20 / 3158.62Source ↗official
SuperCLUE Safetyresponsible_ai#28 / 3145.45Source ↗official
SuperCLUE Safetytraditional_safety#28 / 3165.96Source ↗official