← Models

Model profile

Command R Plus

Coheredeveloper
2024-04-04release date
#202 / 333Safety rank
#161 / 645Freedom rank

Evidence summary

Safety. Command R Plus has an estimated Safety rank of #202; its 90% source-sensitivity interval is #73–#280. Its behavior-only rank is #204; company governance moves the combined estimate to #202. Published Safety evidence spans 12 eval lineages and 7 of 7 components. Its strongest relative result is HELM Safety (simple_safety_tests, #1 of 80); its weakest is Enkrypt AI Safety Leaderboard (bias_attack_non_success_rate, #247 of 248).

Freedom. Command R Plus has an estimated Freedom rank of #161; its 90% source-sensitivity interval is #13–#368. Published Freedom evidence spans 7 eval lineages and 1 of 1 components. Its strongest relative result is AIRBench 2024 Safety Scenarios (safety_scenarios, #2 of 80); its weakest is HELM Safety (simple_safety_tests, #58 of 80).

Compare this model

Only models sharing at least one published sub-eval are listed.

Official and reference links

Safety evals

Rank and direction are specific to the Safety portfolio. Black marks the observed result; the grey dot marks the value implied by the Safety rank. Values stay on each source’s native scale.

Eval / sub-evalRankValueDistributionSource
AgentDojotargeted_attack_success_rate#2 / 150.0445Source ↗official
AgentDojoutility_under_attack#14 / 150.2512Source ↗official
AIRBench 2024 Safety Scenariossafety_scenarios#79 / 800.293Source ↗official
Cisco AI Defense Rolling Single-Turn Leaderboardsingle_turn_attack_success_rate#90 / 10471.39Source ↗official
CRiskEvaldeception_willingness#13 / 1727.04Source ↗official
CRiskEvaldesire_for_resource#10 / 1733.36Source ↗official
CRiskEvalharmful_goal#8 / 1736.67Source ↗official
CRiskEvalimprovement_intent#9 / 1749.73Source ↗official
CRiskEvalmalicious_coordination#13 / 1720.62Source ↗official
CRiskEvalself_preservation#11 / 1737.99Source ↗official
CRiskEvalsituational_awareness#8 / 1764.29Source ↗official
Enkrypt AI Safety Leaderboardbias_attack_non_success_rate#247 / 2481.16Source ↗official
Enkrypt AI Safety Leaderboardcbrn_attack_non_success_rate#181 / 24883.16Source ↗official
Enkrypt AI Safety Leaderboardharmful_attack_non_success_rate#214 / 24839.72Source ↗official
Enkrypt AI Safety Leaderboardtoxicity_attack_non_success_rate#213 / 24688.8Source ↗official
HELM Safetyanthropic_red_team#58 / 800.98Source ↗official
HELM Safetybbq#57 / 800.899Source ↗official
HELM Safetyharmbench#73 / 800.485Source ↗official
HELM Safetysimple_safety_tests#1 / 801Source ↗official
HELM Safetyxstest#58 / 800.938Source ↗official
Large-scale Moral Machine experiment on LLMshuman_choice_distance#13 / 390.8799Source ↗official
Vectara HHEM Factual Consistencyfactual_consistency_rate#26 / 9493.1Source ↗official

Freedom evals

Rank and direction are specific to the Freedom portfolio. Black marks the observed result; the grey dot marks the value implied by the Freedom rank. Values stay on each source’s native scale.

Eval / sub-evalRankValueDistributionSource
AIRBench 2024 Safety Scenariossafety_scenarios#2 / 800.293Source ↗official
Cisco AI Defense Rolling Single-Turn Leaderboardsingle_turn_attack_success_rate#15 / 10471.39Source ↗official
Enkrypt AI Safety Leaderboardcbrn_attack_non_success_rate#68 / 24883.16Source ↗official
Enkrypt AI Safety Leaderboardharmful_attack_non_success_rate#35 / 24839.72Source ↗official
Enkrypt AI Safety Leaderboardtoxicity_attack_non_success_rate#34 / 24688.8Source ↗official
HELM Safetyanthropic_red_team#23 / 800.98Source ↗official
HELM Safetyharmbench#8 / 800.485Source ↗official
HELM Safetysimple_safety_tests#58 / 801Source ↗official
HELM Safetyxstest#58 / 800.938Source ↗official