← Models

Model profile

Claude 3 Haiku

Anthropicdeveloper
2024-03-13release date
#187 / 305overall rank
17eval lineages

Evidence summary

Claude 3 Haiku has an estimated overall rank of #187; its 90% source-sensitivity interval is #74–#219. Its behavior-only rank is #199; company governance moves the combined estimate to #187. Published evidence spans 17 evals and 7 of 7 behavior components. Its strongest relative result is HELM Safety (anthropic_red_team, #1 of 80); its weakest is HELM Safety (bbq, #77 of 80).

Compare this model

Only models sharing at least one published sub-eval are listed.

Official and reference links

Published eval results

Rank is within that sub-eval. Black marks the observed result; the grey dot marks the value implied by the global rank. Values stay on each source’s native scale.

Eval / sub-evalRankValueDistributionSource
AA-Omnisciencehallucination_rate#152 / 3270.8045Source ↗official
AgentDojotargeted_attack_success_rate#5 / 150.0906Source ↗official
AgentDojoutility_under_attack#10 / 150.3339Source ↗official
AgentHarmharm_score#3 / 1211.1Source ↗official
AIRBench 2024 Safety Scenariossafety_scenarios#16 / 800.827Source ↗official
BullshitBench v2clear_pushback_rate#86 / 1060.1Source ↗official
Cisco AI Defense Rolling Single-Turn Leaderboardsingle_turn_attack_success_rate#37 / 10421.14Source ↗official
Confabulationsconfabulation_rate#49 / 5256.93Source ↗official
Contextual MoralChoicehuman_agreement#6 / 220.45Source ↗official
Enkrypt AI Safety Leaderboardbias_attack_non_success_rate#165 / 24112.92Source ↗official
Enkrypt AI Safety Leaderboardcbrn_attack_non_success_rate#40 / 24192.67Source ↗official
Enkrypt AI Safety Leaderboardharmful_attack_non_success_rate#57 / 24187.22Source ↗official
Enkrypt AI Safety Leaderboardtoxicity_attack_non_success_rate#25 / 23999.45Source ↗official
HELM Safetyanthropic_red_team#1 / 801Source ↗official
HELM Safetybbq#77 / 800.625Source ↗official
HELM Safetyharmbench#20 / 800.913Source ↗official
HELM Safetysimple_safety_tests#1 / 801Source ↗official
HELM Safetyxstest#75 / 800.853Source ↗official
Large-scale Moral Machine experiment on LLMshuman_choice_distance#10 / 390.8605Source ↗official
OR-Benchover_refusal_rate#24 / 2596.3Source ↗official
OR-Benchtoxic_acceptance_rate#2 / 250.3Source ↗official
SORRY-Benchavg#5 / 510.08Source ↗official

Values evaluations

Descriptive values and political-framing results are separate from safety/ethics ranks. Each strip shows the evaluation’s observed model range; its endpoint labels state what lower and higher values mean.

UGI Political Values

DimensionValueDistribution
Political Lean-15.7
Government48.5
Diplomacy61.3
Economy47.1
Society60.1