Developer
Mistral AI
39 indexed models; 21 currently meet the evidence threshold for the overall ranking. Together they have results from 58 evaluations.
Company governance evidence
Mistral AI is represented at -1.20 SD relative to the matched Future of Life Institute edition. Provenance: Direct. This company-level evidence contributes 10% of overall model rank.
Models by Mistral AI
Evaluations covering Mistral AI models (58)
AA-Omniscience · AbstentionBench · Adversarial Robustness · AgentDrive Safety Compliance · AgentHarm · AILuminate General Purpose AI Chat · AIRBench 2024 Safety Scenarios · Alignment Leaderboard · AnimalHarmBench · Arena Factuality — Text Arena (factuality-only weighting) · AuAu Authoritarian Response Audit · BlueBench AttaQ-100 · BullshitBench v2 · CASE-Bench · ChineseSafe · Cisco AI Defense Rolling Single-Turn Leaderboard · COMPL-AI AI-Identity Disclosure · COMPL-AI LLM RuLES Multi-Turn Rule Following · COMPL-AI TensorTrust Goal-Hijacking Resistance · Confabulations · Contextual MoralChoice · DSPSafeBench · DystopiaBench · Emergent Collusion · Enkrypt AI Safety Leaderboard · FlagEval Safety and Values · FORTRESS · HalluVerse-M3 Hallucination Recognition · HarmBench · HELM Safety · HUMAINE Trust, Ethics and Safety · JailBench · Large-scale Moral Machine experiment on LLMs · LiveSecBench · MANTA · MASK · Microsoft Phi Safety Panels · MT-JailBench CrescendoX · OR-Bench · PacifAIst · PHARE · Qwen2 Safety Panel · RealityTest — Text AI-Identity Disclosure · RefusalBench · S-Eval · SafeDialBench · SALAD-Bench · Shell · SM-Bench · SORRY-Bench · SpeciEval · StereoTales Harmful Associations · SuperCLUE Safety · TAC · UAVBench safety-critical decision recognition · Vectara HHEM Factual Consistency · VETO Misfired Alignment · XSTest
