← Models

Model profile

GPT 5.4 Mini

OpenAIdeveloper
2026-03-17release date
#72 / 312overall rank
20eval lineages

Evidence summary

GPT 5.4 Mini has an estimated overall rank of #72; its 90% source-sensitivity interval is #27–#218. Its behavior-only rank is #83; company governance moves the combined estimate to #72. Published evidence spans 20 evals and 6 of 7 behavior components. Its strongest relative result is Arena Factuality — Text Arena (factuality-only weighting) (factuality_bt_rating, #1 of 112); its weakest is SM-Bench (overfit, #77 of 79).

Compare this model

Only models sharing at least one published sub-eval are listed.

Official and reference links

Published eval results

Rank is within that sub-eval. Black marks the observed result; the grey dot marks the value implied by the global rank. Values stay on each source’s native scale.

Eval / sub-evalRankValueDistributionSource
AA-Omnisciencehallucination_rate#252 / 3300.8985Source ↗official
Arena Factuality — Text Arena (factuality-only weighting)factuality_bt_rating#1 / 1121497.0Source ↗official
BullshitBench v2clear_pushback_rate#49 / 1060.3167Source ↗official
CAIS Risk Indexagent_red_teaming#23 / 4574.3Source ↗official
CAIS Risk Indexbioweapons_assistance#13 / 5060Source ↗official
CAIS Risk Indexhle_overconfidence#25 / 5155.2Source ↗official
CAIS Risk Indexmachiavelli#40 / 4795.6Source ↗official
CAIS Risk Indexmask#22 / 5315.9Source ↗official
CAIS Risk Indexpolitical_manipulation#32 / 3461.1Source ↗official
CAIS Risk Indextextquests_harm#12 / 5015.4Source ↗official
Cisco AI Defense Rolling Single-Turn Leaderboardsingle_turn_attack_success_rate#26 / 1049.571Source ↗official
DelusionEvaldelusional_prevalence_pct#1 / 1611.2Source ↗official
DelusionEvaldiscourages_harm_prevalence_pct#2 / 1661.3Source ↗official
DelusionEvalfacilitates_harm_prevalence_pct#1 / 160Source ↗official
DelusionEvalrelationship_prevalence_pct#1 / 167.1Source ↗official
DelusionEvalsycophancy_prevalence_pct#2 / 1615.2Source ↗official
DystopiaBenchbasaglia_score#14 / 5049.13Source ↗official
DystopiaBenchbaudrillard_score#12 / 5038.87Source ↗official
DystopiaBenchhuxley_score#11 / 5045.1Source ↗official
DystopiaBenchlaguardia_score#11 / 5052.8Source ↗official
DystopiaBenchorwell_score#12 / 5041.8Source ↗official
DystopiaBenchpetrov_score#9 / 5039.6Source ↗official
Governance Decay under Passive Context Compactiongovernance_retention_score#5 / 759Source ↗official
JuICE Cultural-Error Span Detectionf1#9 / 100.3079Source ↗official
MACHIAVELLIdeception_relative_random_pct#41 / 5095.6Source ↗official
Pander Scoreconversational_absolute_pander_score#12 / 2016.43Source ↗official
Pander Scoreinstructional_absolute_pander_score#10 / 2038.82Source ↗official
RefusalBenchyouden_j#5 / 190.3787Source ↗official
SM-Benchadversarial#66 / 7976.83Source ↗official
SM-Benchambiguous_interpretation#43 / 7984.08Source ↗official
SM-Benchanti_hallucination#68 / 7980.23Source ↗official
SM-Bencheq_boundaries#32 / 7967.28Source ↗official
SM-Benchoverfit#77 / 7915.3Source ↗official
Vectara HHEM Factual Consistencyfactual_consistency_rate#16 / 9494.5Source ↗official
VETO Misfired Alignmentmisfired_alignment_rate_pct#17 / 259.9Source ↗official

Values evaluations

Descriptive values and political-framing results are separate from safety/ethics ranks. Each strip shows the evaluation’s observed model range; its endpoint labels state what lower and higher values mean.

Agent-ValueBench Moral Foundations (MFT08)

Agent-ValueBench HEXACO

DimensionValueDistribution
Openness to experience5.5
Honesty-humility5.3
Extraversion6.1
Agreeableness6.5
Conscientiousness7.3

Agent-ValueBench Schwartz Basic Values (PVQ40)