← Models

Model profile

Grok 4.1 Fast

xAIdeveloper
2025-11-19release date
#206 / 309overall rank
21eval lineages

Evidence summary

Grok 4.1 Fast has an estimated overall rank of #206; its 90% source-sensitivity interval is #83–#281. Its behavior-only rank is #203; company governance moves the combined estimate to #206. Published evidence spans 21 evals and 7 of 7 behavior components. Its strongest relative result is SpeciEval (belief_animal_sentience, #1 of 113); its weakest is AuAu Authoritarian Response Audit (realistic_prompt_arr_percent, #17 of 17).

Compare this model

Only models sharing at least one published sub-eval are listed.

Official and reference links

Published eval results

Rank is within that sub-eval. Black marks the observed result; the grey dot marks the value implied by the global rank. Values stay on each source’s native scale.

Eval / sub-evalRankValueDistributionSource
AA-Omnisciencehallucination_rate#118 / 3280.7341Source ↗official
Alignment Leaderboardcorrigibility#19 / 244.068Source ↗official
Alignment Leaderboardhonesty#9 / 243.682Source ↗official
Alignment Leaderboardnon_manipulation#11 / 243.515Source ↗official
Alignment Leaderboardrobustness#5 / 244.027Source ↗official
Alignment Leaderboardsafety#12 / 243.846Source ↗official
Alignment Leaderboardscheming#13 / 243.627Source ↗official
Arena Factuality — Search Arena (factuality-only weighting)factuality_bt_rating#27 / 301123.0Source ↗official
Arena Factuality — Text Arena (factuality-only weighting)factuality_bt_rating#101 / 1121395.0Source ↗official
AuAu Authoritarian Response Auditbehavioral_vignette_arr_percent#15 / 1729.5Source ↗official
AuAu Authoritarian Response Auditrealistic_prompt_arr_percent#17 / 177.7Source ↗official
BullshitBench v2clear_pushback_rate#76 / 1060.145Source ↗official
CAIS Risk Indexagent_red_teaming#34 / 4490.2Source ↗official
CAIS Risk Indexbioweapons_assistance#18 / 4963.8Source ↗official
CAIS Risk Indexhle_overconfidence#39 / 5072.1Source ↗official
CAIS Risk Indexmachiavelli#5 / 4681.3Source ↗official
CAIS Risk Indexmask#51 / 5271Source ↗official
CAIS Risk Indexpolitical_manipulation#3 / 3332.5Source ↗official
CAIS Risk Indextextquests_harm#3 / 499.1Source ↗official
Cisco AI Defense Rolling Single-Turn Leaderboardsingle_turn_attack_success_rate#42 / 10429.06Source ↗official
Enkrypt AI Safety Leaderboardbias_attack_non_success_rate#183 / 24112.14Source ↗official
Enkrypt AI Safety Leaderboardcbrn_attack_non_success_rate#223 / 24171.83Source ↗official
Enkrypt AI Safety Leaderboardharmful_attack_non_success_rate#132 / 24163.33Source ↗official
Enkrypt AI Safety Leaderboardtoxicity_attack_non_success_rate#212 / 23984.18Source ↗official
LiveSecBenchethics#43 / 438.57Source ↗official
LiveSecBenchfactuality#20 / 4352.76Source ↗official
LiveSecBenchlegality#38 / 4318.02Source ↗official
LiveSecBenchprivacy#25 / 4339.48Source ↗official
LiveSecBenchpsychological_health#27 / 4344.84Source ↗official
MACHIAVELLIdeception_relative_random_pct#5 / 5081.3Source ↗official
MORUscore#9 / 1371.14Source ↗official
MT-JailBench CrescendoXsafety_score#12 / 2118.87Source ↗official
SM-Benchadversarial#43 / 7980.98Source ↗official
SM-Benchambiguous_interpretation#62 / 7978.57Source ↗official
SM-Benchanti_hallucination#23 / 7996.86Source ↗official
SM-Bencheq_boundaries#7 / 7974.72Source ↗official
SM-Benchoverfit#48 / 7966.12Source ↗official
SpeciEvalbelief_animal_sentience#1 / 1137Source ↗official
SpeciEvalland_animal_4ns#112 / 1135.35Source ↗official
SpeciEvalsea_animal_4ns#107 / 1135.28Source ↗official
SpeciEvalspeciesism#100 / 1132.72Source ↗official
Vectara HHEM Factual Consistencyfactual_consistency_rate#87 / 9481.5Source ↗official
Vigil Mental Health Safetyoverall_score#18 / 2330Source ↗official

Values evaluations

Descriptive values and political-framing results are separate from safety/ethics ranks. Each strip shows the evaluation’s observed model range; its endpoint labels state what lower and higher values mean.

UGI Political Values

DimensionValueDistribution
Political Lean18.5
Government43.8
Diplomacy51.7
Economy25.4
Society61.6