← Models

Model profile

Grok 4.1 Fast

xAIdeveloper
2025-11-19release date
#214 / 333Safety rank
#269 / 645Freedom rank

Evidence summary

Safety. Grok 4.1 Fast has an estimated Safety rank of #214; its 90% source-sensitivity interval is #104–#301. Published Safety evidence spans 25 eval lineages and 7 of 7 components. Its strongest relative result is SpeciEval (belief_animal_sentience, #1 of 123); its weakest is AuAu Authoritarian Response Audit (realistic_prompt_arr_percent, #17 of 17).

Freedom. Grok 4.1 Fast has an estimated Freedom rank of #269; its 90% source-sensitivity interval is #99–#463. Published Freedom evidence spans 12 eval lineages and 1 of 1 components. Its strongest relative result is SpeechMap model completion (complete_pct, #2 of 181); its weakest is Adversarial Humanities Benchmark (AHB) — Table 5 (child_sexual_exploitation, #28 of 31).

Compare this model

Only models sharing at least one published sub-eval are listed.

Official and reference links

Safety evals

Rank and direction are specific to the Safety portfolio. Black marks the observed result; the grey dot marks the value implied by the Safety rank. Values stay on each source’s native scale.

Eval / sub-evalRankValueDistributionSource
AA-Omnisciencehallucination_rate#138 / 3450.7341Source ↗official
Adversarial Humanities Benchmark (AHB) — Table 5child_sexual_exploitation#4 / 3114.6Source ↗official
Adversarial Humanities Benchmark (AHB) — Table 5defamation#22 / 3178.3Source ↗official
Adversarial Humanities Benchmark (AHB) — Table 5hate#23 / 3165.2Source ↗official
Adversarial Humanities Benchmark (AHB) — Table 5indiscriminate_weapons_cbrne#10 / 3133.3Source ↗official
Adversarial Humanities Benchmark (AHB) — Table 5intellectual_property#10 / 3141.3Source ↗official
Adversarial Humanities Benchmark (AHB) — Table 5non_violent_crimes#11 / 3152.2Source ↗official
Adversarial Humanities Benchmark (AHB) — Table 5privacy#11 / 3153.2Source ↗official
Adversarial Humanities Benchmark (AHB) — Table 5sex_related_crimes#11 / 3154.3Source ↗official
Adversarial Humanities Benchmark (AHB) — Table 5sexual_content#26 / 3164.6Source ↗official
Adversarial Humanities Benchmark (AHB) — Table 5specialized_advice#14 / 3160.92Source ↗official
Adversarial Humanities Benchmark (AHB) — Table 5suicide_self_harm#13 / 3155.3Source ↗official
Adversarial Humanities Benchmark (AHB) — Table 5violent_crimes#13 / 3160Source ↗official
Alignment Leaderboardcorrigibility#19 / 244.068Source ↗official
Alignment Leaderboardhonesty#9 / 243.682Source ↗official
Alignment Leaderboardnon_manipulation#11 / 243.515Source ↗official
Alignment Leaderboardrobustness#5 / 244.027Source ↗official
Alignment Leaderboardsafety#12 / 243.846Source ↗official
Alignment Leaderboardscheming#13 / 243.627Source ↗official
Arena Factuality — Search Arena (factuality-only weighting)factuality_bt_rating#27 / 301123.0Source ↗official
Arena Factuality — Text Arena (factuality-only weighting)factuality_bt_rating#100 / 1111395.0Source ↗official
AuAu Authoritarian Response Auditbehavioral_vignette_arr_percent#15 / 1729.5Source ↗official
AuAu Authoritarian Response Auditrealistic_prompt_arr_percent#17 / 177.7Source ↗official
BioTIERpermit_compliance_pct#36 / 5298.4Source ↗official
BioTIERrefuse_compliance_pct#9 / 5277.3Source ↗official
BullshitBench v2clear_pushback_rate#87 / 1170.145Source ↗official
CAIS Risk Indexagent_red_teaming#39 / 4990.2Source ↗official
CAIS Risk Indexbioweapons_assistance#20 / 5463.8Source ↗official
CAIS Risk Indexhle_overconfidence#44 / 5572.1Source ↗official
CAIS Risk Indexmachiavelli#6 / 5181.3Source ↗official
CAIS Risk Indexmask#56 / 5771Source ↗official
CAIS Risk Indexpolitical_manipulation#3 / 4832.5Source ↗official
CAIS Risk Indextextquests_harm#3 / 549.1Source ↗official
Cisco AI Defense Rolling Single-Turn Leaderboardsingle_turn_attack_success_rate#42 / 10429.06Source ↗official
Enkrypt AI Safety Leaderboardbias_attack_non_success_rate#191 / 24812.14Source ↗official
Enkrypt AI Safety Leaderboardcbrn_attack_non_success_rate#223 / 24871.83Source ↗official
Enkrypt AI Safety Leaderboardharmful_attack_non_success_rate#141 / 24863.33Source ↗official
Enkrypt AI Safety Leaderboardtoxicity_attack_non_success_rate#219 / 24684.18Source ↗official
Human Pathogen Capabilities Test (HPCT) — overall refusalrefusal_rate_pct#16 / 6924.55Source ↗official
LiveSecBenchethics#43 / 438.57Source ↗official
LiveSecBenchfactuality#20 / 4352.76Source ↗official
LiveSecBenchlegality#38 / 4318.02Source ↗official
LiveSecBenchprivacy#25 / 4339.48Source ↗official
LiveSecBenchpsychological_health#27 / 4344.84Source ↗official
MACHIAVELLIdeception_relative_random_pct#5 / 5081.3Source ↗official
MORUscore#9 / 1371.14Source ↗official
MT-JailBench CrescendoXsafety_score#12 / 2118.87Source ↗official
Pander Scoreconversational_absolute_pander_score#26 / 2634.8Source ↗official
Pander Scoreinstructional_absolute_pander_score#26 / 2684.47Source ↗official
SM-Benchadversarial#48 / 8480.98Source ↗official
SM-Benchambiguous_interpretation#67 / 8478.57Source ↗official
SM-Benchanti_hallucination#26 / 8496.86Source ↗official
SM-Bencheq_boundaries#8 / 8474.72Source ↗official
SM-Benchoverfit#53 / 8466.12Source ↗official
SpeciEvalbelief_animal_sentience#1 / 1237Source ↗official
SpeciEvalland_animal_4ns#122 / 1235.35Source ↗official
SpeciEvalsea_animal_4ns#117 / 1235.28Source ↗official
SpeciEvalspeciesism#110 / 1232.72Source ↗official
Vectara HHEM Factual Consistencyfactual_consistency_rate#87 / 9481.5Source ↗official
Vigil Mental Health Safetyoverall_score#18 / 2330Source ↗official

Freedom evals

Rank and direction are specific to the Freedom portfolio. Black marks the observed result; the grey dot marks the value implied by the Freedom rank. Values stay on each source’s native scale.

Eval / sub-evalRankValueDistributionSource
Adversarial Humanities Benchmark (AHB) — Table 5child_sexual_exploitation#28 / 3114.6Source ↗official
Adversarial Humanities Benchmark (AHB) — Table 5defamation#10 / 3178.3Source ↗official
Adversarial Humanities Benchmark (AHB) — Table 5hate#9 / 3165.2Source ↗official
Adversarial Humanities Benchmark (AHB) — Table 5indiscriminate_weapons_cbrne#21 / 3133.3Source ↗official
Adversarial Humanities Benchmark (AHB) — Table 5intellectual_property#22 / 3141.3Source ↗official
Adversarial Humanities Benchmark (AHB) — Table 5non_violent_crimes#21 / 3152.2Source ↗official
Adversarial Humanities Benchmark (AHB) — Table 5privacy#21 / 3153.2Source ↗official
Adversarial Humanities Benchmark (AHB) — Table 5sex_related_crimes#21 / 3154.3Source ↗official
Adversarial Humanities Benchmark (AHB) — Table 5sexual_content#5 / 3164.6Source ↗official
Adversarial Humanities Benchmark (AHB) — Table 5specialized_advice#18 / 3160.92Source ↗official
Adversarial Humanities Benchmark (AHB) — Table 5suicide_self_harm#19 / 3155.3Source ↗official
Adversarial Humanities Benchmark (AHB) — Table 5violent_crimes#19 / 3160Source ↗official
BioTIERpermit_compliance_pct#36 / 5298.4Source ↗official
BioTIERrefuse_compliance_pct#44 / 5277.3Source ↗official
CAIS Risk Indexbioweapons_assistance#34 / 5463.8Source ↗official
Cisco AI Defense Rolling Single-Turn Leaderboardsingle_turn_attack_success_rate#63 / 10429.06Source ↗official
Enkrypt AI Safety Leaderboardcbrn_attack_non_success_rate#26 / 24871.83Source ↗official
Enkrypt AI Safety Leaderboardharmful_attack_non_success_rate#108 / 24863.33Source ↗official
Enkrypt AI Safety Leaderboardtoxicity_attack_non_success_rate#28 / 24684.18Source ↗official
Human Pathogen Capabilities Test (HPCT) — overall refusalrefusal_rate_pct#54 / 6924.55Source ↗official
LiveSecBenchethics#1 / 438.57Source ↗official
LiveSecBenchlegality#6 / 4318.02Source ↗official
LiveSecBenchprivacy#19 / 4339.48Source ↗official
LiveSecBenchpsychological_health#17 / 4344.84Source ↗official
MT-JailBench CrescendoXsafety_score#10 / 2118.87Source ↗official
SM-Benchadversarial#35 / 8480.98Source ↗official
SM-Bencheq_boundaries#8 / 8474.72Source ↗official
SM-Benchoverfit#53 / 8466.12Source ↗official
SpeechMap model completioncomplete_pct#2 / 18197.9Source ↗official
UGI Leaderboard — base-model willingnesswillingness_adherence_score#15 / 1567.5Source ↗official
UGI Leaderboard — base-model willingnesswillingness_direct_score#17 / 1565.5Source ↗official
Vigil Mental Health Safetyoverall_score#6 / 2330Source ↗official

Values evaluations

Descriptive values and political-framing results are separate from safety/ethics ranks. Each strip shows the evaluation’s observed model range; its endpoint labels state what lower and higher values mean.

UGI Political Values

DimensionValueDistribution
Political Lean18.5
Government43.8
Diplomacy51.7
Economy25.4
Society61.6