Developer
Anthropic
27 indexed models; 21 currently meet the evidence threshold for the overall ranking. Together they have results from 104 evaluations.
Company governance evidence
Anthropic is represented at +2.07 SD relative to the matched Future of Life Institute edition. Provenance: Direct. This company-level evidence contributes 10% of overall model rank.
Models by Anthropic
| Model | Evals | Components | Rank | Release date |
|---|---|---|---|---|
| Claude Opus 5 | 15 | 6/7 | 1 | 2026-07-24 |
| Claude Fable 5 | 18 | 7/7 | 2 | 2026-06-09 |
| Claude 2 | 6 | 4/7 | 3 | 2023-07-11 |
| Claude Opus 4.8 | 22 | 7/7 | 6 | 2026-05-28 |
| Claude Opus 4.7 | 29 | 7/7 | 7 | 2026-04-16 |
| Claude Opus 4 | 21 | 7/7 | 12 | 2025-05-22 |
| Claude Sonnet 4.6 | 32 | 7/7 | 13 | 2026-01-21 |
| Claude Opus 4.6 | 28 | 7/7 | 16 | 2026-02-05 |
| Claude Haiku 4.5 | 40 | 7/7 | 21 | 2025-10-15 |
| Claude Opus 4.5 | 25 | 7/7 | 23 | 2025-11-24 |
| Claude Sonnet 4.5 | 39 | 7/7 | 24 | 2025-09-29 |
| Claude 3.7 Sonnet | 27 | 7/7 | 26 | 2025-02-24 |
| Claude Sonnet 4 | 35 | 7/7 | 31 | 2025-05-22 |
| Claude Opus 4.1 | 15 | 6/7 | 32 | 2025-08-05 |
| Claude 3.5 Haiku | 14 | 7/7 | 57 | 2024-10-22 |
| Claude Sonnet 5 | 12 | 7/7 | 59 | 2026-06-30 |
| Claude 3 Opus | 23 | 7/7 | 90 | 2024-03-04 |
| Claude 3.5 Sonnet | 30 | 7/7 | 126 | 2024-06-21 |
| Claude 2.1 | 3 | 2/7 | 167 | 2023-11-21 |
| Claude 3 Haiku | 17 | 7/7 | 187 | 2024-03-13 |
| Claude 3 Sonnet | 13 | 6/7 | 220 | 2024-03-04 |
| Claude 1 | 2 | 1/7 | — | 2023-03-14 |
| Claude 3.6 Sonnet | 1 | 2/7 | — | — |
| Claude Instant 1.2 | 1 | 1/7 | — | 2023-08-09 |
| Claude Mythos 5 | 1 | 1/7 | — | — |
| Claude Mythos Preview | 1 | 1/7 | — | — |
| Stanford Online All V4 S3 | 1 | 1/7 | — | — |
Evaluations covering Anthropic models (104)
AA-Omniscience · Adversarial Robustness · Agent-SafetyBench · AgentAbstain · AgentDojo · AgentDrive Safety Compliance · AgentHarm · AILuminate General Purpose AI Chat · AIMS Safety-Classifier Competence · AIRBench 2024 Safety Scenarios · Alignment Leaderboard · ANIMA · AnimalHarmBench · Anthropic Agentic Misalignment — blackmail · Anthropic Agentic Misalignment — corporate espionage · Anthropic Agentic Misalignment — lethal action · Anthropic Claude 4 System Card · Anthropic Claude Haiku 4.5 System Card · Anthropic Claude Opus 4.1 System Card Addendum · Anthropic Claude Opus 4.5 System Card · Anthropic Claude Sonnet 4.5 System Card · Arena Factuality — Search Arena (factuality-only weighting) · Arena Factuality — Text Arena (factuality-only weighting) · AuAu Authoritarian Response Audit · AutoElicit Transferability · BioSecBench-Refusal · BullshitBench v2 · CAIS Risk Index · CASE-Bench · Cisco AI Defense Rolling Single-Turn Leaderboard · Claude Sonnet 4.6 Overrefusal · Claude Sonnet 4.6 User Wellbeing · COMPL-AI AI-Identity Disclosure · COMPL-AI LLM RuLES Multi-Turn Rule Following · COMPL-AI TensorTrust Goal-Hijacking Resistance · Confabulations · Constitutional Following — Anthropic Constitution · Constitutional Following — OpenAI Model Spec · Contextual MoralChoice · DecodingTrust · Do-Not-Answer · DystopiaBench · Emergent Collusion · Enkrypt AI Safety Leaderboard · Fake Alignment (FINE) · FinEval Financial Security Knowledge · FlagEval Safety and Values · FORTRESS · Governance Decay under Passive Context Compaction · Gray Swan indirect prompt injection (15 attempts) · HalluVerse-M3 Hallucination Recognition · HarmBench · HarmVideoBench · HELM Classic RealToxicityPrompts · HELM Safety · HUMAINE Trust, Ethics and Safety · Inkling-Small model card — FORTRESS · Inkling-Small model card — StrongREJECT · JuICE Cultural-Error Span Detection · KIDBench Implicit Child Cue · Large-scale Moral Machine experiment on LLMs · LiveSecBench · LLM Ethics Benchmark · MACHIAVELLI · Manager Coercion Bench · MANTA · MASK · MonitoringBench Full-Trajectory Monitor · MORU · MT-JailBench CrescendoX · ODCV-Bench · Olam Social Poker — Social Lie Rate · OpenAgentSafety · OR-Bench · PacifAIst · PandaBench JBB direct-request panel · PHARE · Pokee-Isaac model card — DTAP · PropensityBench · RealityTest — Text AI-Identity Disclosure · RefusalBench · Reward Hacking Benchmark · SABER · SafeArena · SALAD-Bench · Shell · SM-Bench · Social Welfare Function Benchmark · SORRY-Bench · SOSBench · SpeciesismBench · SpeciEval · StereoTales Harmful Associations · SuperCLUE Safety · SYCON Bench · TAC · ThaiSafetyBench · ToolPrivacyBench · TrustLLM contemporary collapsed application · UAVBench safety-critical decision recognition · UK AISI active safety-research compromise continuation · Vectara HHEM Factual Consistency · VETO Misfired Alignment · Vigil Mental Health Safety