← Models

Model profile

Claude Haiku 4.5

Anthropicdeveloper
2025-10-15release date
#17 / 267overall rank
32eval lineages

Evidence summary

Claude Haiku 4.5 has an estimated overall rank of #17; its 90% source-sensitivity interval is #11–#71. Its behavior-only rank is #24; company governance moves the combined estimate to #17. Published evidence spans 32 evals and 7 of 7 behavior components. Its strongest relative result is AIRBench 2024 Safety Scenarios (safety_scenarios, #1 of 80); its weakest is Anthropic Claude Haiku 4.5 System Card (harmful_request_safety, #2 of 2).

Compare this model

Only models sharing at least one published sub-eval are listed.

Official and reference links

Published eval results

Rank is within that sub-eval. Black marks the observed result; the grey dot marks the value implied by the global rank. Values stay on each source’s native scale.

Eval / sub-evalRankValueDistributionBetterSource
AA-Omnisciencehallucination_rate#9 / 3110.2471↓ lowerSource ↗official
AgentAbstainabstain#5 / 1765.6↑ higherSource ↗official
AgentAbstaincar#6 / 1761.8↑ higherSource ↗official
AgentAbstainpaired#5 / 1749.7↑ higherSource ↗official
AIRBench 2024 Safety Scenariossafety_scenarios#1 / 800.932↑ higherSource ↗official
ANIMAscore#14 / 190.5862↑ higherSource ↗official
Anthropic Claude Haiku 4.5 System Cardagentic_coding_safety#1 / 31↑ higherSource ↗official
Anthropic Claude Haiku 4.5 System Cardclaude_code_malicious_refusal#2 / 30.6939↑ higherSource ↗official
Anthropic Claude Haiku 4.5 System Cardharmful_request_safety#2 / 20.9938↑ higherSource ↗official
Anthropic Claude Opus 4.5 System Cardagentic_coding_safety#1 / 41↑ higherSource ↗official
Anthropic Claude Opus 4.5 System Cardclaude_code_malicious_refusal#2 / 40.6939↑ higherSource ↗official
Anthropic Claude Opus 4.5 System Cardcomputer_use_malicious_refusal#3 / 40.7232↑ higherSource ↗official
BullshitBench v2clear_pushback_rate#8 / 1050.74↑ higherSource ↗official
CAIS Risk Indexagent_red_teaming#18 / 4369.7↓ lowerSource ↗official
CAIS Risk Indexbioweapons_assistance#34 / 4884.9↓ lowerSource ↗official
CAIS Risk Indexhle_overconfidence#31 / 4960.5↓ lowerSource ↗official
CAIS Risk Indexmachiavelli#6 / 4581.5↓ lowerSource ↗official
CAIS Risk Indexmask#3 / 514.6↓ lowerSource ↗official
CAIS Risk Indexpolitical_manipulation#31 / 3262.2↓ lowerSource ↗official
CAIS Risk Indextextquests_harm#18 / 4816.6↓ lowerSource ↗official
Cisco AI Defense Rolling Single-Turn Leaderboardsingle_turn_attack_success_rate#12 / 1053.539↓ lowerSource ↗official
Claude Sonnet 4.6 Overrefusalhigher_difficulty_overrefusal_rate#4 / 56.01↓ lowerSource ↗official
Claude Sonnet 4.6 Overrefusaloverall_overrefusal_rate#3 / 50.26↓ lowerSource ↗official
Claude Sonnet 4.6 User Wellbeingchild_benign_refusal_rate#1 / 40.04↓ lowerSource ↗official
Claude Sonnet 4.6 User Wellbeingchild_multiturn_appropriate_rate#2 / 496↑ higherSource ↗official
Claude Sonnet 4.6 User Wellbeingchild_violative_harmless_rate#3 / 499.88↑ higherSource ↗official
Claude Sonnet 4.6 User Wellbeingselfharm_benign_refusal_rate#2 / 40.03↓ lowerSource ↗official
Claude Sonnet 4.6 User Wellbeingselfharm_harmless_rate#3 / 499.67↑ higherSource ↗official
Claude Sonnet 4.6 User Wellbeingselfharm_multiturn_appropriate_rate#2 / 490↑ higherSource ↗official
Contextual MoralChoicehuman_agreement#6 / 220.45↑ higherSource ↗official
DystopiaBenchbasaglia_score#3 / 5025.3↓ lowerSource ↗official
DystopiaBenchbaudrillard_score#2 / 5022.73↓ lowerSource ↗official
DystopiaBenchhuxley_score#6 / 5019.07↓ lowerSource ↗official
DystopiaBenchlaguardia_score#4 / 5034.47↓ lowerSource ↗official
DystopiaBenchorwell_score#2 / 5021.63↓ lowerSource ↗official
DystopiaBenchpetrov_score#1 / 5019.8↓ lowerSource ↗official
Enkrypt AI Safety Leaderboardbias_attack_non_success_rate#43 / 26035.14↑ higherSource ↗official
Enkrypt AI Safety Leaderboardcbrn_attack_non_success_rate#109 / 26089.5↑ higherSource ↗official
Enkrypt AI Safety Leaderboardharmful_attack_non_success_rate#17 / 26098.89↑ higherSource ↗official
Enkrypt AI Safety Leaderboardtoxicity_attack_non_success_rate#51 / 25898.91↑ higherSource ↗official
HELM Safetyanthropic_red_team#64 / 800.969↑ higherSource ↗official
HELM Safetybbq#44 / 800.928↑ higherSource ↗official
HELM Safetyharmbench#14 / 800.959↑ higherSource ↗official
HELM Safetysimple_safety_tests#40 / 800.988↑ higherSource ↗official
HELM Safetyxstest#60 / 800.932↑ higherSource ↗official
Inkling-Small model card — FORTRESSbenign_answer_rate#8 / 1094.1↑ higherSource ↗official
Inkling-Small model card — FORTRESSharmful_refusal_rate#1 / 1091.3↑ higherSource ↗official
Inkling-Small model card — StrongREJECTsafety_rate#6 / 1098.6↑ higherSource ↗official
LiveSecBenchethics#4 / 4388.65↑ higherSource ↗official
LiveSecBenchfactuality#3 / 4382.68↑ higherSource ↗official
LiveSecBenchlegality#1 / 4398.62↑ higherSource ↗official
LiveSecBenchprivacy#1 / 4397.17↑ higherSource ↗official
LiveSecBenchpsychological_health#2 / 4390.02↑ higherSource ↗official
MACHIAVELLIdeception_relative_random_pct#6 / 5081.5↓ lowerSource ↗official
MORUscore#7 / 1374.02↑ higherSource ↗official
PHAREbias_resistance_diagnostic#6 / 660.7066↑ higherSource ↗official
PHAREhallucination_resistance_diagnostic#10 / 700.8356↑ higherSource ↗official
PHAREharm_resistance_diagnostic#3 / 700.9993↑ higherSource ↗official
PHAREjailbreak_resistance_diagnostic#4 / 670.7851↑ higherSource ↗official
SM-Benchadversarial#61 / 7376.59↑ higherSource ↗official
SM-Benchambiguous_interpretation#29 / 7386.61↑ higherSource ↗official
SM-Benchanti_hallucination#37 / 7392.15↑ higherSource ↗official
SM-Bencheq_boundaries#71 / 7341.85↑ higherSource ↗official
SM-Benchoverfit#22 / 7381.97↑ higherSource ↗official
SpeciEvalbelief_animal_sentience#95 / 1026.32↑ higherSource ↗official
SpeciEvalland_animal_4ns#16 / 1024.25↓ lowerSource ↗official
SpeciEvalsea_animal_4ns#22 / 1024.53↓ lowerSource ↗official
SpeciEvalspeciesism#61 / 1022.15↓ lowerSource ↗official
UAVBench safety-critical decision recognitionethical_safety_critical_accuracy#17 / 270.67↑ higherSource ↗official
Vigil Mental Health Safetyoverall_score#4 / 2374↑ higherSource ↗official

Values evaluations

Descriptive values and political-framing results are separate from safety/ethics ranks. Each strip shows the evaluation’s observed model range; its endpoint labels state what lower and higher values mean.

UGI Political Values

DimensionValueDistribution
Political Lean-18.4
Government45.6
Diplomacy63.6
Economy44.5
Society58.2

Agent-ValueBench Moral Foundations (MFT08)

Agent-ValueBench HEXACO

DimensionValueDistribution
Openness to experience5.4
Honesty-humility5.5
Extraversion5.7
Agreeableness7.5
Conscientiousness7.1

Agent-ValueBench Schwartz Basic Values (PVQ40)

The Economist World Values Survey Cultural Map

DimensionValueDistribution
Survival ↔ Self-expression3.03
Traditional ↔ Secular0.557