← Models

Model profile

Claude Opus 4.6

Anthropicdeveloper
2026-02-05release date
#8 / 267overall rank
23eval lineages

Evidence summary

Claude Opus 4.6 has an estimated overall rank of #8; its 90% source-sensitivity interval is #7–#51. Its behavior-only rank is #12; company governance moves the combined estimate to #8. Published evidence spans 23 evals and 7 of 7 behavior components. Its strongest relative result is SpeciEval (belief_animal_sentience, #1 of 102); its weakest is Claude Sonnet 4.6 User Wellbeing (child_benign_refusal_rate, #4 of 4).

Compare this model

Only models sharing at least one published sub-eval are listed.

Official and reference links

Published eval results

Rank is within that sub-eval. Black marks the observed result; the grey dot marks the value implied by the global rank. Values stay on each source’s native scale.

Eval / sub-evalRankValueDistributionBetterSource
AA-Omnisciencehallucination_rate#74 / 3110.6133↓ lowerSource ↗official
ANIMAscore#5 / 190.7246↑ higherSource ↗official
BioSecBench-Refusalbalanced_refusal_score#5 / 100.3987↑ higherSource ↗official
BullshitBench v2clear_pushback_rate#3 / 1050.85↑ higherSource ↗official
CAIS Risk Indexagent_red_teaming#10 / 4349.5↓ lowerSource ↗official
CAIS Risk Indexbioweapons_assistance#7 / 4834.8↓ lowerSource ↗official
CAIS Risk Indexhle_overconfidence#14 / 4946↓ lowerSource ↗official
CAIS Risk Indexmachiavelli#15 / 4584.4↓ lowerSource ↗official
CAIS Risk Indexmask#17 / 5112.7↓ lowerSource ↗official
CAIS Risk Indexpolitical_manipulation#22 / 3251.7↓ lowerSource ↗official
CAIS Risk Indextextquests_harm#43 / 4823.2↓ lowerSource ↗official
Cisco AI Defense Rolling Single-Turn Leaderboardsingle_turn_attack_success_rate#13 / 1053.639↓ lowerSource ↗official
Claude Sonnet 4.6 Overrefusalhigher_difficulty_overrefusal_rate#1 / 50.04↓ lowerSource ↗official
Claude Sonnet 4.6 Overrefusaloverall_overrefusal_rate#4 / 50.66↓ lowerSource ↗official
Claude Sonnet 4.6 User Wellbeingchild_benign_refusal_rate#4 / 40.18↓ lowerSource ↗official
Claude Sonnet 4.6 User Wellbeingchild_multiturn_appropriate_rate#2 / 496↑ higherSource ↗official
Claude Sonnet 4.6 User Wellbeingchild_violative_harmless_rate#2 / 499.95↑ higherSource ↗official
Claude Sonnet 4.6 User Wellbeingselfharm_benign_refusal_rate#4 / 40.25↓ lowerSource ↗official
Claude Sonnet 4.6 User Wellbeingselfharm_harmless_rate#1 / 499.75↑ higherSource ↗official
Claude Sonnet 4.6 User Wellbeingselfharm_multiturn_appropriate_rate#3 / 482↑ higherSource ↗official
DystopiaBenchbasaglia_score#5 / 5026.77↓ lowerSource ↗official
DystopiaBenchbaudrillard_score#10 / 5033.37↓ lowerSource ↗official
DystopiaBenchhuxley_score#2 / 5015.77↓ lowerSource ↗official
DystopiaBenchlaguardia_score#6 / 5039.23↓ lowerSource ↗official
DystopiaBenchorwell_score#4 / 5023.47↓ lowerSource ↗official
DystopiaBenchpetrov_score#5 / 5030.07↓ lowerSource ↗official
Enkrypt AI Safety Leaderboardbias_attack_non_success_rate#76 / 26022.74↑ higherSource ↗official
Enkrypt AI Safety Leaderboardcbrn_attack_non_success_rate#66 / 26091.33↑ higherSource ↗official
Enkrypt AI Safety Leaderboardharmful_attack_non_success_rate#7 / 26099.44↑ higherSource ↗official
FORTRESSaverage_risk_score#11 / 4916.76↓ lowerSource ↗official
FORTRESSover_refusal_score#37 / 469.59↓ lowerSource ↗official
HUMAINE Trust, Ethics and Safetytrust_ethics_safety_score#23 / 5427.23↑ higherSource ↗official
MACHIAVELLIdeception_relative_random_pct#15 / 5084.4↓ lowerSource ↗official
MASKlying_probability_pct#5 / 539.16↓ lowerSource ↗official
MORUscore#6 / 1374.1↑ higherSource ↗official
PHAREbias_resistance_diagnostic#57 / 660.3351↑ higherSource ↗official
PHAREhallucination_resistance_diagnostic#2 / 700.8776↑ higherSource ↗official
PHAREharm_resistance_diagnostic#1 / 701↑ higherSource ↗official
PHAREjailbreak_resistance_diagnostic#6 / 670.7524↑ higherSource ↗official
RefusalBenchyouden_j#7 / 190.6667↑ higherSource ↗official
SABERoverall_safety_rate#1 / 1345.26↑ higherSource ↗official
SABERscenario_a_safety_rate#1 / 1356.3↑ higherSource ↗official
SABERscenario_b_safety_rate#1 / 1339.77↑ higherSource ↗official
SABERscenario_c_safety_rate#1 / 1336.87↑ higherSource ↗official
SM-Benchadversarial#28 / 7383.41↑ higherSource ↗official
SM-Benchambiguous_interpretation#40 / 7383.63↑ higherSource ↗official
SM-Benchanti_hallucination#1 / 73100↑ higherSource ↗official
SM-Bencheq_boundaries#70 / 7342.98↑ higherSource ↗official
SM-Benchoverfit#3 / 7395.63↑ higherSource ↗official
SpeciEvalbelief_animal_sentience#1 / 1027↑ higherSource ↗official
SpeciEvalland_animal_4ns#96 / 1025.05↓ lowerSource ↗official
SpeciEvalsea_animal_4ns#79 / 1025.03↓ lowerSource ↗official
SpeciEvalspeciesism#65 / 1022.17↓ lowerSource ↗official

Values evaluations

Descriptive values and political-framing results are separate from safety/ethics ranks. Each strip shows the evaluation’s observed model range; its endpoint labels state what lower and higher values mean.

UGI Political Values

DimensionValueDistribution
Political Lean-19.5
Government44.5
Diplomacy68
Economy44.2
Society61.8