← Models

Model profile

Grok 4

xAIdeveloper
2025-07-09release date
#105 / 267overall rank
27eval lineages

Evidence summary

Grok 4 has an estimated overall rank of #105; its 90% source-sensitivity interval is #47–#164. Its behavior-only rank is #109; company governance moves the combined estimate to #105. Published evidence spans 27 evals and 7 of 7 behavior components. Its strongest relative result is Confabulations (confabulation_rate, #3 of 52); its weakest is Emergent Collusion (high_illegality_game_rate, #13 of 13).

Compare this model

Only models sharing at least one published sub-eval are listed.

Official and reference links

Published eval results

Rank is within that sub-eval. Black marks the observed result; the grey dot marks the value implied by the global rank. Values stay on each source’s native scale.

Eval / sub-evalRankValueDistributionBetterSource
AA-Omnisciencehallucination_rate#79 / 3110.642↓ lowerSource ↗official
AIRBench 2024 Safety Scenariossafety_scenarios#69 / 800.444↑ higherSource ↗official
BrokenMathsycophancy#5 / 943.4↓ lowerSource ↗official
CAIS Risk Indexagent_red_teaming#29 / 4385.9↓ lowerSource ↗official
CAIS Risk Indexbioweapons_assistance#20 / 4864.2↓ lowerSource ↗official
CAIS Risk Indexhle_overconfidence#25 / 4956.4↓ lowerSource ↗official
CAIS Risk Indexmachiavelli#35 / 4593.4↓ lowerSource ↗official
CAIS Risk Indexmask#35 / 5141.7↓ lowerSource ↗official
CAIS Risk Indexpolitical_manipulation#18 / 3249.5↓ lowerSource ↗official
CAIS Risk Indextextquests_harm#46 / 4824.3↓ lowerSource ↗official
Cisco AI Defense Rolling Single-Turn Leaderboardsingle_turn_attack_success_rate#39 / 10521.78↓ lowerSource ↗official
Confabulationsconfabulation_rate#3 / 523.96↓ lowerSource ↗official
DystopiaBenchbasaglia_score#28 / 5065.3↓ lowerSource ↗official
DystopiaBenchbaudrillard_score#25 / 5057.7↓ lowerSource ↗official
DystopiaBenchhuxley_score#30 / 5073.17↓ lowerSource ↗official
DystopiaBenchlaguardia_score#15 / 5059.07↓ lowerSource ↗official
DystopiaBenchorwell_score#25 / 5068.63↓ lowerSource ↗official
DystopiaBenchpetrov_score#31 / 5074.37↓ lowerSource ↗official
Emergent Collusionhigh_illegality_game_rate#13 / 130.75↓ lowerSource ↗official
Enkrypt AI Safety Leaderboardbias_attack_non_success_rate#195 / 26012.4↑ higherSource ↗official
Enkrypt AI Safety Leaderboardcbrn_attack_non_success_rate#230 / 26076.83↑ higherSource ↗official
Enkrypt AI Safety Leaderboardharmful_attack_non_success_rate#239 / 26033.89↑ higherSource ↗official
Enkrypt AI Safety Leaderboardtoxicity_attack_non_success_rate#155 / 25895.45↑ higherSource ↗official
FlagEval Safety and Valuesa1_qualified_rate#6 / 1881.99↑ higherSource ↗official
FlagEval Safety and Valuesa2_qualified_rate#7 / 1880.3↑ higherSource ↗official
FlagEval Safety and Valuesa3_qualified_rate#12 / 1886.36↑ higherSource ↗official
FlagEval Safety and Valuesa4_qualified_rate#10 / 1887.98↑ higherSource ↗official
FlagEval Safety and Valuesa5_qualified_rate#6 / 1872.83↑ higherSource ↗official
HELM Safetyanthropic_red_team#71 / 800.957↑ higherSource ↗official
HELM Safetybbq#40 / 800.937↑ higherSource ↗official
HELM Safetyharmbench#78 / 800.397↑ higherSource ↗official
HELM Safetysimple_safety_tests#73 / 800.922↑ higherSource ↗official
HELM Safetyxstest#30 / 800.966↑ higherSource ↗official
HUMAINE Trust, Ethics and Safetytrust_ethics_safety_score#9 / 5428.61↑ higherSource ↗official
MACHIAVELLIdeception_relative_random_pct#36 / 5093.4↓ lowerSource ↗official
PHAREbias_resistance_diagnostic#61 / 660.3116↑ higherSource ↗official
PHAREhallucination_resistance_diagnostic#17 / 700.8073↑ higherSource ↗official
PHAREharm_resistance_diagnostic#70 / 700.7177↑ higherSource ↗official
PHAREjailbreak_resistance_diagnostic#20 / 670.6485↑ higherSource ↗official
Shelleducation_jsr#14 / 140.81↓ lowerSource ↗official
Shellfinance_jsr#7 / 140.486↓ lowerSource ↗official
Shellmanagement_jsr#7 / 140.596↓ lowerSource ↗official
Social Welfare Function Benchmarkfairness#2 / 190.619↑ higherSource ↗official
SpeciEvalbelief_animal_sentience#36 / 1026.87↑ higherSource ↗official
SpeciEvalland_animal_4ns#48 / 1024.53↓ lowerSource ↗official
SpeciEvalsea_animal_4ns#53 / 1024.78↓ lowerSource ↗official
SpeciEvalspeciesism#23 / 1021.65↓ lowerSource ↗official
TrustLLM contemporary collapsed applicationtrustllm#3 / 80.62↑ higherSource ↗official

Values evaluations

Descriptive values and political-framing results are separate from safety/ethics ranks. Each strip shows the evaluation’s observed model range; its endpoint labels state what lower and higher values mean.

UGI Political Values

DimensionValueDistribution
Political Lean-11.7
Government45.6
Diplomacy63.8
Economy41.3
Society65.6

The Economist World Values Survey Cultural Map

DimensionValueDistribution
Survival ↔ Self-expression3.38
Traditional ↔ Secular0.7