← Models

Model profile

Claude Opus 4.7

Anthropicdeveloper
2026-04-16release date
#10 / 309overall rank
29eval lineages
2discovery sources

Evidence summary

Claude Opus 4.7 has an estimated overall rank of #10; its 90% source-sensitivity interval is #6–#61. Its behavior-only rank is #15; company governance moves the combined estimate to #10. Published evidence spans 29 evals and 7 of 7 behavior components. Its strongest relative result is Enkrypt AI Safety Leaderboard (harmful_attack_non_success_rate, #6 of 241); its weakest is MORU (score, #12 of 13).

Compare this model

Only models sharing at least one published sub-eval are listed.

Official and reference links

Published eval results

Rank is within that sub-eval. Black marks the observed result; the grey dot marks the value implied by the global rank. Values stay on each source’s native scale.

Eval / sub-evalRankValueDistributionSource
AA-Omnisciencehallucination_rate#44 / 3280.4229Source ↗official
AgentAbstainabstain#1 / 1779Source ↗official
AgentAbstaincar#1 / 1777.6Source ↗official
AgentAbstainpaired#2 / 1759.4Source ↗official
ANIMAscore#5 / 220.735Source ↗official
Arena Factuality — Search Arena (factuality-only weighting)factuality_bt_rating#23 / 301171.0Source ↗official
Arena Factuality — Text Arena (factuality-only weighting)factuality_bt_rating#45 / 1121442.0Source ↗official
BioSecBench-Refusalbalanced_refusal_score#3 / 100.4259Source ↗official
BullshitBench v2clear_pushback_rate#6 / 1060.785Source ↗official
CAIS Risk Indexagent_red_teaming#6 / 4444.5Source ↗official
CAIS Risk Indexbioweapons_assistance#3 / 4919.2Source ↗official
CAIS Risk Indexhle_overconfidence#5 / 5028.1Source ↗official
CAIS Risk Indexmachiavelli#19 / 4685.7Source ↗official
CAIS Risk Indexmask#24 / 5217.1Source ↗official
CAIS Risk Indexpolitical_manipulation#18 / 3348.2Source ↗official
CAIS Risk Indextextquests_harm#39 / 4921.8Source ↗official
Cisco AI Defense Rolling Single-Turn Leaderboardsingle_turn_attack_success_rate#7 / 1042.642Source ↗official
DystopiaBenchbasaglia_score#4 / 5026.03Source ↗official
DystopiaBenchbaudrillard_score#5 / 5027.5Source ↗official
DystopiaBenchhuxley_score#4 / 5018.33Source ↗official
DystopiaBenchlaguardia_score#2 / 5030.07Source ↗official
DystopiaBenchorwell_score#3 / 5022.73Source ↗official
DystopiaBenchpetrov_score#4 / 5026.77Source ↗official
Enkrypt AI Safety Leaderboardbias_attack_non_success_rate#64 / 24123.51Source ↗official
Enkrypt AI Safety Leaderboardcbrn_attack_non_success_rate#9 / 24196.33Source ↗official
Enkrypt AI Safety Leaderboardharmful_attack_non_success_rate#6 / 24199.44Source ↗official
Enkrypt AI Safety Leaderboardtoxicity_attack_non_success_rate#50 / 23998.55Source ↗official
HarmVideoBenchharmful_video_safety_recognition_reasoning#4 / 190.803Source ↗official
HUMAINE Trust, Ethics and Safetytrust_ethics_safety_score#24 / 5427.19Source ↗official
JuICE Cultural-Error Span Detectionf1#4 / 100.4503Source ↗official
MACHIAVELLIdeception_relative_random_pct#19 / 5085.7Source ↗official
MANTAAWMS#1 / 70.579Source ↗official
MANTAAWVS#1 / 70.76Source ↗official
MORUscore#12 / 1367.2Source ↗official
ODCV-Benchaverage_severity#1 / 120.0125Source ↗official
ODCV-Benchmisalignment_rate#1 / 120Source ↗official
RefusalBenchyouden_j#7 / 190.234Source ↗official
SM-Benchadversarial#40 / 7981.46Source ↗official
SM-Benchambiguous_interpretation#19 / 7989.29Source ↗official
SM-Benchanti_hallucination#20 / 7997.38Source ↗official
SM-Bencheq_boundaries#57 / 7957.87Source ↗official
SM-Benchoverfit#9 / 7991.8Source ↗official
SpeciEvalbelief_animal_sentience#37 / 1136.88Source ↗official
SpeciEvalland_animal_4ns#60 / 1134.57Source ↗official
SpeciEvalsea_animal_4ns#36 / 1134.65Source ↗official
SpeciEvalspeciesism#54 / 1131.98Source ↗official
ToolPrivacyBenchprivate_mt_poi#2 / 920.31Source ↗official
ToolPrivacyBenchpublic_mt_poi#4 / 917.83Source ↗official
UK AISI active safety-research compromise continuationactive_compromise_continuation_rate_pct#2 / 50.8Source ↗official
Vectara HHEM Factual Consistencyfactual_consistency_rate#71 / 9488Source ↗official
VETO Misfired Alignmentmisfired_alignment_rate_pct#20 / 2510.7Source ↗official

Values evaluations

Descriptive values and political-framing results are separate from safety/ethics ranks. Each strip shows the evaluation’s observed model range; its endpoint labels state what lower and higher values mean.

UGI Political Values

DimensionValueDistribution
Political Lean-14.3
Government46
Diplomacy62.2
Economy45.4
Society58.9

The Economist World Values Survey Cultural Map

DimensionValueDistribution
Survival ↔ Self-expression3.13
Traditional ↔ Secular-0.217