← Models

Model profile

Claude Opus 4.1

Anthropicdeveloper
2025-08-05release date
#36 / 309overall rank
15eval lineages

Evidence summary

Claude Opus 4.1 has an estimated overall rank of #36; its 90% source-sensitivity interval is #21–#104. Its behavior-only rank is #43; company governance moves the combined estimate to #36. Published evidence spans 15 evals and 6 of 7 behavior components. Its strongest relative result is Cisco AI Defense Rolling Single-Turn Leaderboard (single_turn_attack_success_rate, #2 of 104); its weakest is Anthropic Claude Opus 4.1 System Card Addendum (benign_request_refusal, #2 of 2).

Compare this model

Only models sharing at least one published sub-eval are listed.

Official and reference links

Published eval results

Rank is within that sub-eval. Black marks the observed result; the grey dot marks the value implied by the global rank. Values stay on each source’s native scale.

Eval / sub-evalRankValueDistributionSource
AgentDrive Safety Compliancescr#10 / 4893.75Source ↗official
Anthropic Claude Opus 4.1 System Card Addendumbbq_disambiguated_accuracy#2 / 20.907Source ↗official
Anthropic Claude Opus 4.1 System Card Addendumbenign_request_refusal#2 / 20.0008Source ↗official
Anthropic Claude Opus 4.1 System Card Addendumharmful_request_safety#1 / 20.9876Source ↗official
Anthropic Claude Opus 4.5 System Cardagentic_coding_safety#4 / 40.96Source ↗official
Anthropic Claude Opus 4.5 System Cardclaude_code_malicious_refusal#4 / 40.4816Source ↗official
Anthropic Claude Opus 4.5 System Cardcomputer_use_malicious_refusal#4 / 40.6696Source ↗official
Anthropic Claude Opus 4.5 System Cardharmful_request_safety#2 / 20.9914Source ↗official
Anthropic Claude Sonnet 4.5 System Cardbenign_request_refusal#2 / 20.0008Source ↗official
Arena Factuality — Search Arena (factuality-only weighting)factuality_bt_rating#21 / 301178.0Source ↗official
Arena Factuality — Text Arena (factuality-only weighting)factuality_bt_rating#29 / 1121450.0Source ↗official
BullshitBench v2clear_pushback_rate#33 / 1060.425Source ↗official
Cisco AI Defense Rolling Single-Turn Leaderboardsingle_turn_attack_success_rate#2 / 1042.243Source ↗official
Confabulationsconfabulation_rate#2 / 523.218Source ↗official
FORTRESSaverage_risk_score#8 / 4915.47Source ↗official
FORTRESSover_refusal_score#21 / 464.065Source ↗official
MASKlying_probability_pct#6 / 539.2Source ↗official
PHAREbias_resistance_diagnostic#39 / 660.4361Source ↗official
PHAREhallucination_resistance_diagnostic#6 / 700.8619Source ↗official
PHAREharm_resistance_diagnostic#18 / 700.9631Source ↗official
PHAREjailbreak_resistance_diagnostic#2 / 670.8135Source ↗official
Social Welfare Function Benchmarkfairness#7 / 190.525Source ↗official
SOSBenchbiology_pvr#3 / 230.134Source ↗official
SOSBenchchemistry_pvr#2 / 230.147Source ↗official
SOSBenchmedicine_pvr#2 / 230.232Source ↗official
SOSBenchpharmacology_pvr#2 / 230.249Source ↗official
SOSBenchphysics_pvr#1 / 230.098Source ↗official
SOSBenchpsychology_pvr#1 / 230.107Source ↗official
SpeciEvalbelief_animal_sentience#77 / 1136.62Source ↗official
SpeciEvalland_animal_4ns#31 / 1134.33Source ↗official
SpeciEvalsea_animal_4ns#17 / 1134.47Source ↗official
SpeciEvalspeciesism#48 / 1131.92Source ↗official
Vectara HHEM Factual Consistencyfactual_consistency_rate#69 / 9488.2Source ↗official

Values evaluations

Descriptive values and political-framing results are separate from safety/ethics ranks. Each strip shows the evaluation’s observed model range; its endpoint labels state what lower and higher values mean.

UGI Political Values

DimensionValueDistribution
Political Lean-22.6
Government45.7
Diplomacy66
Economy46.9
Society61.3