Developer
56 indexed models; 32 currently meet the evidence threshold for the overall ranking. Together they have results from 96 evaluations.
Company governance evidence
Google is represented at +0.80 SD relative to the matched Future of Life Institute edition. Provenance: Direct. This company-level evidence contributes 10% of overall model rank.
Models by Google
Evaluations covering Google models (96)
AA-Omniscience · AbstentionBench · Adversarial Robustness · Agent-SafetyBench · AgentAbstain · AgentDojo · AgentDrive Safety Compliance · AILuminate General Purpose AI Chat · AIMS Safety-Classifier Competence · AIRBench 2024 Safety Scenarios · Alignment Leaderboard · ANIMA · AnimalHarmBench · Anthropic Agentic Misalignment — blackmail · Anthropic Agentic Misalignment — corporate espionage · Anthropic Agentic Misalignment — lethal action · Arena Factuality — Search Arena (factuality-only weighting) · Arena Factuality — Text Arena (factuality-only weighting) · AuAu Authoritarian Response Audit · BioSecBench-Refusal · BrokenMath · BullshitBench v2 · CAIS Risk Index · ChineseSafe · Cisco AI Defense Rolling Single-Turn Leaderboard · COMPL-AI AI-Identity Disclosure · COMPL-AI LLM RuLES Multi-Turn Rule Following · COMPL-AI TensorTrust Goal-Hijacking Resistance · Confabulations · Constitutional Following — Anthropic Constitution · Constitutional Following — OpenAI Model Spec · DecodingTrust · DSPSafeBench · DystopiaBench · Emergent Collusion · Enkrypt AI Safety Leaderboard · FinEval Financial Security Knowledge · FlagEval Safety and Values · FORTRESS · Google Gemini 2.5 Flash Model Card · Google Gemini 2.5 Flash-Lite Model Card · Governance Decay under Passive Context Compaction · Gray Swan indirect prompt injection (15 attempts) · HalluVerse-M3 Hallucination Recognition · HarmBench · HarmVideoBench · HELM Classic RealToxicityPrompts · HELM Safety · HUMAINE Trust, Ethics and Safety · IndoBias-Pairs — parity-aware culturally grounded bias · Inkling-Small model card — FORTRESS · Inkling-Small model card — StrongREJECT · JuICE Cultural-Error Span Detection · KIDBench Implicit Child Cue · Large-scale Moral Machine experiment on LLMs · LiveSecBench · LLM Ethics Benchmark · MACHIAVELLI · Manager Coercion Bench · MANTA · MASK · Microsoft Phi Safety Panels · MORU · MT-JailBench CrescendoX · MuPPET Contextual Privacy · ODCV-Bench · Olam Social Poker — Social Lie Rate · OR-Bench · PacifAIst · PandaBench JBB direct-request panel · PHARE · Pokee-Isaac model card — DTAP · PropensityBench · RealityTest — Text AI-Identity Disclosure · RefusalBench · Reward Hacking Benchmark · S-Eval · SafetyBench · SALAD-Bench · Shell · SM-Bench · Social Welfare Function Benchmark · SORRY-Bench · SOSBench · SpeciesismBench · SpeciEval · StereoTales Harmful Associations · SYCON Bench · TAC · ThaiSafetyBench · ToolPrivacyBench · TrustLLM contemporary collapsed application · UAVBench safety-critical decision recognition · Vectara HHEM Factual Consistency · VETO Misfired Alignment · Vigil Mental Health Safety
