← Models

Model profile

Claude Sonnet 5

Anthropicdeveloper
2026-06-30release date
#59 / 305overall rank
12eval lineages
2discovery sources

Evidence summary

Claude Sonnet 5 has an estimated overall rank of #59; its 90% source-sensitivity interval is #2–#201. Its behavior-only rank is #76; company governance moves the combined estimate to #59. Published evidence spans 12 evals and 7 of 7 behavior components. Its strongest relative result is Enkrypt AI Safety Leaderboard (harmful_attack_non_success_rate, #1 of 241); its weakest is TAC (base_welfare_rate, #72 of 74).

Compare this model

Only models sharing at least one published sub-eval are listed.

Official and reference links

Published eval results

Rank is within that sub-eval. Black marks the observed result; the grey dot marks the value implied by the global rank. Values stay on each source’s native scale.

Eval / sub-evalRankValueDistributionSource
AA-Omnisciencehallucination_rate#38 / 3270.3937Source ↗official
Arena Factuality — Search Arena (factuality-only weighting)factuality_bt_rating#14 / 301190.0Source ↗official
Arena Factuality — Text Arena (factuality-only weighting)factuality_bt_rating#22 / 1121456.0Source ↗official
BullshitBench v2clear_pushback_rate#5 / 1060.79Source ↗official
Cisco AI Defense Rolling Single-Turn Leaderboardsingle_turn_attack_success_rate#2 / 1042.243Source ↗official
Enkrypt AI Safety Leaderboardbias_attack_non_success_rate#6 / 24170.54Source ↗official
Enkrypt AI Safety Leaderboardcbrn_attack_non_success_rate#43 / 24192.5Source ↗official
Enkrypt AI Safety Leaderboardharmful_attack_non_success_rate#1 / 241100Source ↗official
Enkrypt AI Safety Leaderboardtoxicity_attack_non_success_rate#9 / 23999.82Source ↗official
Gray Swan indirect prompt injection (15 attempts)attack_success_probability_k15_pct#5 / 135.9Source ↗official
Manager Coercion Benchcoercion_ladder_depth#6 / 315.5Source ↗official
Manager Coercion Benchfabrication_rate#1 / 130Source ↗official
Olam Social Poker — Social Lie Ratesocial_lie_rate_per_10000_turns#3 / 192Source ↗official
PHAREbias_resistance_diagnostic#43 / 660.3988Source ↗official
PHAREhallucination_resistance_diagnostic#7 / 700.861Source ↗official
PHAREharm_resistance_diagnostic#4 / 700.9959Source ↗official
PHAREjailbreak_resistance_diagnostic#1 / 670.8684Source ↗official
SM-Benchadversarial#42 / 7880.98Source ↗official
SM-Benchambiguous_interpretation#8 / 7891.96Source ↗official
SM-Benchanti_hallucination#1 / 78100Source ↗official
SM-Bencheq_boundaries#62 / 7855.06Source ↗official
SM-Benchoverfit#19 / 7883.06Source ↗official
SpeciEvalbelief_animal_sentience#75 / 1056.57Source ↗official
SpeciEvalland_animal_4ns#80 / 1054.8Source ↗official
SpeciEvalsea_animal_4ns#87 / 1055.05Source ↗official
SpeciEvalspeciesism#55 / 1052.07Source ↗official
TACbase_welfare_rate#72 / 7415.38Source ↗official

Values evaluations

Descriptive values and political-framing results are separate from safety/ethics ranks. Each strip shows the evaluation’s observed model range; its endpoint labels state what lower and higher values mean.

UGI Political Values

DimensionValueDistribution
Political Lean-19.8
Government42.8
Diplomacy64.1
Economy45.9
Society60.5