← Models

Model profile

Claude Opus 4.7

Anthropicdeveloper
2026-04-16release date
#5 / 267overall rank
25eval lineages
2discovery sources

Evidence summary

Claude Opus 4.7 has an estimated overall rank of #5; its 90% source-sensitivity interval is #3–#50. Its behavior-only rank is #6; company governance moves the combined estimate to #5. Published evidence spans 25 evals and 7 of 7 behavior components. Its strongest relative result is Enkrypt AI Safety Leaderboard (harmful_attack_non_success_rate, #7 of 260); its weakest is MORU (score, #12 of 13).

Compare this model

Only models sharing at least one published sub-eval are listed.

Official and reference links

Published eval results

Rank is within that sub-eval. Black marks the observed result; the grey dot marks the value implied by the global rank. Values stay on each source’s native scale.

Eval / sub-evalRankValueDistributionBetterSource
AA-Omnisciencehallucination_rate#27 / 3110.3618↓ lowerSource ↗official
AgentAbstainabstain#1 / 1779↑ higherSource ↗official
AgentAbstaincar#1 / 1777.6↑ higherSource ↗official
AgentAbstainpaired#2 / 1759.4↑ higherSource ↗official
ANIMAscore#4 / 190.735↑ higherSource ↗official
BioSecBench-Refusalbalanced_refusal_score#3 / 100.4259↑ higherSource ↗official
BullshitBench v2clear_pushback_rate#6 / 1050.785↑ higherSource ↗official
CAIS Risk Indexagent_red_teaming#5 / 4344.5↓ lowerSource ↗official
CAIS Risk Indexbioweapons_assistance#3 / 4819.2↓ lowerSource ↗official
CAIS Risk Indexhle_overconfidence#4 / 4928.1↓ lowerSource ↗official
CAIS Risk Indexmachiavelli#19 / 4585.7↓ lowerSource ↗official
CAIS Risk Indexmask#23 / 5117.1↓ lowerSource ↗official
CAIS Risk Indexpolitical_manipulation#16 / 3248.2↓ lowerSource ↗official
CAIS Risk Indextextquests_harm#39 / 4821.8↓ lowerSource ↗official
Cisco AI Defense Rolling Single-Turn Leaderboardsingle_turn_attack_success_rate#7 / 1052.642↓ lowerSource ↗official
DystopiaBenchbasaglia_score#4 / 5026.03↓ lowerSource ↗official
DystopiaBenchbaudrillard_score#5 / 5027.5↓ lowerSource ↗official
DystopiaBenchhuxley_score#4 / 5018.33↓ lowerSource ↗official
DystopiaBenchlaguardia_score#2 / 5030.07↓ lowerSource ↗official
DystopiaBenchorwell_score#3 / 5022.73↓ lowerSource ↗official
DystopiaBenchpetrov_score#4 / 5026.77↓ lowerSource ↗official
Enkrypt AI Safety Leaderboardbias_attack_non_success_rate#68 / 26023.51↑ higherSource ↗official
Enkrypt AI Safety Leaderboardcbrn_attack_non_success_rate#12 / 26096.33↑ higherSource ↗official
Enkrypt AI Safety Leaderboardharmful_attack_non_success_rate#7 / 26099.44↑ higherSource ↗official
Enkrypt AI Safety Leaderboardtoxicity_attack_non_success_rate#56 / 25898.55↑ higherSource ↗official
HUMAINE Trust, Ethics and Safetytrust_ethics_safety_score#24 / 5427.19↑ higherSource ↗official
MACHIAVELLIdeception_relative_random_pct#19 / 5085.7↓ lowerSource ↗official
MANTAAWMS#1 / 70.579↑ higherSource ↗official
MANTAAWVS#1 / 70.76↑ higherSource ↗official
MORUscore#12 / 1367.2↑ higherSource ↗official
ODCV-Benchaverage_severity#1 / 120.0125↓ lowerSource ↗official
ODCV-Benchmisalignment_rate#1 / 120↓ lowerSource ↗official
RefusalBenchyouden_j#12 / 190.234↑ higherSource ↗official
SM-Benchadversarial#37 / 7381.46↑ higherSource ↗official
SM-Benchambiguous_interpretation#15 / 7389.29↑ higherSource ↗official
SM-Benchanti_hallucination#17 / 7397.38↑ higherSource ↗official
SM-Bencheq_boundaries#52 / 7357.87↑ higherSource ↗official
SM-Benchoverfit#8 / 7391.8↑ higherSource ↗official
SpeciEvalbelief_animal_sentience#33 / 1026.88↑ higherSource ↗official
SpeciEvalland_animal_4ns#53 / 1024.57↓ lowerSource ↗official
SpeciEvalsea_animal_4ns#29 / 1024.65↓ lowerSource ↗official
SpeciEvalspeciesism#49 / 1021.98↓ lowerSource ↗official
ToolPrivacyBenchprivate_mt_poi#2 / 920.31↓ lowerSource ↗official
ToolPrivacyBenchpublic_mt_poi#4 / 917.83↓ lowerSource ↗official
UK AISI active safety-research compromise continuationactive_compromise_continuation_rate_pct#2 / 50.8↓ lowerSource ↗official
VETO Misfired Alignmentmisfired_alignment_rate_pct#20 / 2510.7↓ lowerSource ↗official

Values evaluations

Descriptive values and political-framing results are separate from safety/ethics ranks. Each strip shows the evaluation’s observed model range; its endpoint labels state what lower and higher values mean.

UGI Political Values

DimensionValueDistribution
Political Lean-14.3
Government46
Diplomacy62.2
Economy45.4
Society58.9

The Economist World Values Survey Cultural Map

DimensionValueDistribution
Survival ↔ Self-expression3.13
Traditional ↔ Secular-0.217