Developer
Meta
36 indexed models; 20 currently meet the evidence threshold for the overall ranking. Together they have results from 82 evaluations.
Company governance evidence
Meta is represented at -1.24 SD relative to the matched Future of Life Institute edition. Provenance: Direct. This company-level evidence contributes 10% of overall model rank.
Models by Meta
Evaluations covering Meta models (82)
AA-Omniscience · AbstentionBench · Adversarial Robustness · Agent-SafetyBench · AgentDojo · AgentDrive Safety Compliance · AgentHarm · AILuminate General Purpose AI Chat · AIMS Safety-Classifier Competence · AIRBench 2024 Safety Scenarios · Alignment Leaderboard · AnimalHarmBench · Anthropic Agentic Misalignment — blackmail · Anthropic Agentic Misalignment — corporate espionage · Arena Factuality — Text Arena (factuality-only weighting) · BlueBench AttaQ-100 · BullshitBench v2 · CAIS Risk Index · CASE-Bench · ChineseSafe · Cisco AI Defense Rolling Single-Turn Leaderboard · COMPL-AI AI-Identity Disclosure · COMPL-AI LLM RuLES Multi-Turn Rule Following · COMPL-AI TensorTrust Goal-Hijacking Resistance · Confabulations · Contextual MoralChoice · DecodingTrust · Do-Not-Answer · DSPSafeBench · DystopiaBench · Enkrypt AI Safety Leaderboard · FinEval 6.0 Safety Awareness · FORTRESS · Gray Swan indirect prompt injection (15 attempts) · HalluVerse-M3 Hallucination Recognition · HarmBench · HELM Classic RealToxicityPrompts · HELM Safety · HUMAINE Trust, Ethics and Safety · IndoBias-Pairs — parity-aware culturally grounded bias · JailBench · JuICE Cultural-Error Span Detection · KIDBench Implicit Child Cue · Large-scale Moral Machine experiment on LLMs · LiveSecBench · LLM Ethics Benchmark · MACHIAVELLI · Manager Coercion Bench · MANTA · MASK · Microsoft Phi Safety Panels · MT-JailBench CrescendoX · MuPPET Contextual Privacy · ODCV-Bench · Olam Social Poker — Social Lie Rate · Open LLM Safety Index · OR-Bench · PandaBench JBB direct-request panel · PHARE · PropensityBench · RealityTest — Text AI-Identity Disclosure · RefusalBench · S-Eval · SafeArena · SafetyBench · SALAD-Bench · Shell · SM-Bench · SORRY-Bench · SOSBench · SpeciesismBench · SpeciEval · SYCON Bench · TAC · ThaiSafetyBench · TrustLLM contemporary collapsed application · TukaBench · UAVBench safety-critical decision recognition · Vectara HHEM Factual Consistency · VETO Misfired Alignment · Vigil Mental Health Safety · XSTest