← Models

Model profile

Claude 3.5 Haiku

Anthropicdeveloper
2024-10-22release date
#46 / 267overall rank
14eval lineages

Evidence summary

Claude 3.5 Haiku has an estimated overall rank of #46; its 90% source-sensitivity interval is #5–#150. Its behavior-only rank is #62; company governance moves the combined estimate to #46. Published evidence spans 14 evals and 7 of 7 behavior components. Its strongest relative result is Enkrypt AI Safety Leaderboard (cbrn_attack_non_success_rate, #10 of 260); its weakest is Confabulations (confabulation_rate, #51 of 52).

Compare this model

Only models sharing at least one published sub-eval are listed.

Official and reference links

Published eval results

Rank is within that sub-eval. Black marks the observed result; the grey dot marks the value implied by the global rank. Values stay on each source’s native scale.

Eval / sub-evalRankValueDistributionBetterSource
AA-Omnisciencehallucination_rate#35 / 3110.4154↓ lowerSource ↗official
Agent-SafetyBenchcompromise_availability#11 / 1626.4↑ higherSource ↗official
Agent-SafetyBenchharmful_vulnerable_code#2 / 1660.8↑ higherSource ↗official
Agent-SafetyBenchleak_sensitive_information#3 / 1647.2↑ higherSource ↗official
Agent-SafetyBenchphysical_harm#4 / 1645.6↑ higherSource ↗official
Agent-SafetyBenchproduce_unsafe_information#1 / 16100↑ higherSource ↗official
Agent-SafetyBenchproperty_loss#5 / 1646↑ higherSource ↗official
Agent-SafetyBenchspread_unsafe_information#2 / 1633.6↑ higherSource ↗official
Agent-SafetyBenchviolate_law_ethics#3 / 1641.2↑ higherSource ↗official
AILuminate General Purpose AI Chatchild_sexual_exploitation#2 / 321.8↓ lowerSource ↗official
AILuminate General Purpose AI Chatdefamation#3 / 322.5↓ lowerSource ↗official
AILuminate General Purpose AI Chathate#3 / 321↓ lowerSource ↗official
AILuminate General Purpose AI Chatindiscriminate_weapons_cbrne#3 / 323.3↓ lowerSource ↗official
AILuminate General Purpose AI Chatintellectual_property#3 / 322↓ lowerSource ↗official
AILuminate General Purpose AI Chatnon_violent_crimes#2 / 323↓ lowerSource ↗official
AILuminate General Purpose AI Chatprivacy#3 / 322.2↓ lowerSource ↗official
AILuminate General Purpose AI Chatsex_related_crimes#2 / 323↓ lowerSource ↗official
AILuminate General Purpose AI Chatsexual_content#3 / 322.503↓ lowerSource ↗official
AILuminate General Purpose AI Chatspecialized_advice#3 / 323.715↓ lowerSource ↗official
AILuminate General Purpose AI Chatsuicide_self_harm#3 / 322.8↓ lowerSource ↗official
AILuminate General Purpose AI Chatviolent_crimes#2 / 322.9↓ lowerSource ↗official
AnimalHarmBenchscore#6 / 100.02↑ higherSource ↗official
Anthropic Agentic Misalignment — blackmailmisaligned_action_rate_pct#3 / 1610↓ lowerSource ↗official
Anthropic Agentic Misalignment — corporate espionagemisaligned_action_rate_pct#4 / 1619↓ lowerSource ↗official
Anthropic Claude Haiku 4.5 System Cardagentic_coding_safety#1 / 31↑ higherSource ↗official
Anthropic Claude Haiku 4.5 System Cardclaude_code_malicious_refusal#1 / 30.7↑ higherSource ↗official
Anthropic Claude Haiku 4.5 System Cardharmful_request_safety#1 / 20.9972↑ higherSource ↗official
BullshitBench v2clear_pushback_rate#21 / 1050.5↑ higherSource ↗official
Cisco AI Defense Rolling Single-Turn Leaderboardsingle_turn_attack_success_rate#36 / 10520.74↓ lowerSource ↗official
Confabulationsconfabulation_rate#51 / 5265.84↓ lowerSource ↗official
Enkrypt AI Safety Leaderboardbias_attack_non_success_rate#11 / 26056.59↑ higherSource ↗official
Enkrypt AI Safety Leaderboardcbrn_attack_non_success_rate#10 / 26097↑ higherSource ↗official
Enkrypt AI Safety Leaderboardharmful_attack_non_success_rate#37 / 26094.44↑ higherSource ↗official
Enkrypt AI Safety Leaderboardtoxicity_attack_non_success_rate#34 / 25899.36↑ higherSource ↗official
FORTRESSaverage_risk_score#26 / 4930.41↓ lowerSource ↗official
FORTRESSover_refusal_score#43 / 4613.16↓ lowerSource ↗official
Large-scale Moral Machine experiment on LLMshuman_choice_distance#24 / 391.118↓ lowerSource ↗official
PandaBench JBB direct-request panelsafety_rate#14 / 460.99↑ higherSource ↗official
PHAREbias_resistance_diagnostic#49 / 660.3808↑ higherSource ↗official
PHAREhallucination_resistance_diagnostic#27 / 700.7804↑ higherSource ↗official
PHAREharm_resistance_diagnostic#25 / 700.9536↑ higherSource ↗official
PHAREjailbreak_resistance_diagnostic#21 / 670.6482↑ higherSource ↗official

Values evaluations

Descriptive values and political-framing results are separate from safety/ethics ranks. Each strip shows the evaluation’s observed model range; its endpoint labels state what lower and higher values mean.

ValueCompass

DimensionValueDistribution
Universalism70.1
Self-direction50.6
Care / Harm55.9
Fairness / Cheating54.7
Ethical90.4