← Models

Model profile

Claude 3 Haiku

Anthropicdeveloper
2024-03-13release date
#161 / 267overall rank
17eval lineages

Evidence summary

Claude 3 Haiku has an estimated overall rank of #161; its 90% source-sensitivity interval is #68–#197. Its behavior-only rank is #174; company governance moves the combined estimate to #161. Published evidence spans 17 evals and 7 of 7 behavior components. Its strongest relative result is HELM Safety (anthropic_red_team, #1 of 80); its weakest is HELM Safety (bbq, #77 of 80).

Compare this model

Only models sharing at least one published sub-eval are listed.

Official and reference links

Published eval results

Rank is within that sub-eval. Black marks the observed result; the grey dot marks the value implied by the global rank. Values stay on each source’s native scale.

Eval / sub-evalRankValueDistributionBetterSource
AA-Omnisciencehallucination_rate#134 / 3110.7825↓ lowerSource ↗official
AgentDojotargeted_attack_success_rate#5 / 150.0906↓ lowerSource ↗official
AgentDojoutility_under_attack#10 / 150.3339↑ higherSource ↗official
AgentHarmharm_score#3 / 1211.1↓ lowerSource ↗official
AIRBench 2024 Safety Scenariossafety_scenarios#16 / 800.827↑ higherSource ↗official
BullshitBench v2clear_pushback_rate#85 / 1050.1↑ higherSource ↗official
Cisco AI Defense Rolling Single-Turn Leaderboardsingle_turn_attack_success_rate#37 / 10521.14↓ lowerSource ↗official
Confabulationsconfabulation_rate#49 / 5256.93↓ lowerSource ↗official
Contextual MoralChoicehuman_agreement#6 / 220.45↑ higherSource ↗official
Enkrypt AI Safety Leaderboardbias_attack_non_success_rate#181 / 26012.92↑ higherSource ↗official
Enkrypt AI Safety Leaderboardcbrn_attack_non_success_rate#46 / 26092.67↑ higherSource ↗official
Enkrypt AI Safety Leaderboardharmful_attack_non_success_rate#66 / 26087.22↑ higherSource ↗official
Enkrypt AI Safety Leaderboardtoxicity_attack_non_success_rate#31 / 25899.45↑ higherSource ↗official
HELM Safetyanthropic_red_team#1 / 801↑ higherSource ↗official
HELM Safetybbq#77 / 800.625↑ higherSource ↗official
HELM Safetyharmbench#20 / 800.913↑ higherSource ↗official
HELM Safetysimple_safety_tests#1 / 801↑ higherSource ↗official
HELM Safetyxstest#75 / 800.853↑ higherSource ↗official
Large-scale Moral Machine experiment on LLMshuman_choice_distance#10 / 390.8605↓ lowerSource ↗official
OR-Benchover_refusal_rate#24 / 2596.3↓ lowerSource ↗official
OR-Benchtoxic_acceptance_rate#2 / 250.3↓ lowerSource ↗official
SORRY-Benchavg#5 / 510.08↓ lowerSource ↗official

Values evaluations

Descriptive values and political-framing results are separate from safety/ethics ranks. Each strip shows the evaluation’s observed model range; its endpoint labels state what lower and higher values mean.

UGI Political Values

DimensionValueDistribution
Political Lean-15.7
Government48.5
Diplomacy61.3
Economy47.1
Society60.1