← Models

Model profile

GPT 5.4

OpenAIdeveloper
2026-03-05release date
#11 / 305overall rank
27eval lineages

Evidence summary

GPT 5.4 has an estimated overall rank of #11; its 90% source-sensitivity interval is #7–#65. Its behavior-only rank is #13; company governance moves the combined estimate to #11. Published evidence spans 27 evals and 7 of 7 behavior components. Its strongest relative result is Enkrypt AI Safety Leaderboard (harmful_attack_non_success_rate, #1 of 241); its weakest is CAIS Risk Index (textquests_harm, #48 of 49).

Compare this model

Only models sharing at least one published sub-eval are listed.

Official and reference links

Published eval results

Rank is within that sub-eval. Black marks the observed result; the grey dot marks the value implied by the global rank. Values stay on each source’s native scale.

Eval / sub-evalRankValueDistributionSource
AA-Omnisciencehallucination_rate#175 / 3270.8258Source ↗official
AgentAbstainabstain#3 / 1767.8Source ↗official
AgentAbstaincar#5 / 1764Source ↗official
AgentAbstainpaired#7 / 1748.7Source ↗official
AIMS Safety-Classifier Competenceaverage_harmful_f1#1 / 110.815Source ↗official
Arena Factuality — Search Arena (factuality-only weighting)factuality_bt_rating#3 / 301262.0Source ↗official
Arena Factuality — Text Arena (factuality-only weighting)factuality_bt_rating#2 / 1121496.0Source ↗official
BioSecBench-Refusalbalanced_refusal_score#4 / 100.4254Source ↗official
BullshitBench v2clear_pushback_rate#29 / 1060.45Source ↗official
CAIS Risk Indexagent_red_teaming#14 / 4459.7Source ↗official
CAIS Risk Indexbioweapons_assistance#12 / 4957.1Source ↗official
CAIS Risk Indexhle_overconfidence#7 / 5040.9Source ↗official
CAIS Risk Indexmachiavelli#34 / 4692.9Source ↗official
CAIS Risk Indexmask#12 / 529.7Source ↗official
CAIS Risk Indexpolitical_manipulation#29 / 3359.4Source ↗official
CAIS Risk Indextextquests_harm#48 / 4924.5Source ↗official
Cisco AI Defense Rolling Single-Turn Leaderboardsingle_turn_attack_success_rate#8 / 1042.742Source ↗official
DystopiaBenchbasaglia_score#13 / 5047.3Source ↗official
DystopiaBenchbaudrillard_score#12 / 5038.87Source ↗official
DystopiaBenchhuxley_score#13 / 5046.2Source ↗official
DystopiaBenchlaguardia_score#10 / 5046.57Source ↗official
DystopiaBenchorwell_score#13 / 5044.4Source ↗official
DystopiaBenchpetrov_score#7 / 5037.4Source ↗official
Enkrypt AI Safety Leaderboardbias_attack_non_success_rate#165 / 24112.92Source ↗official
Enkrypt AI Safety Leaderboardcbrn_attack_non_success_rate#181 / 24183.33Source ↗official
Enkrypt AI Safety Leaderboardharmful_attack_non_success_rate#1 / 241100Source ↗official
Enkrypt AI Safety Leaderboardtoxicity_attack_non_success_rate#94 / 23997.27Source ↗official
GPT-5.6 system cardconnectors_injection_resistance#5 / 70.998Source ↗official
GPT-5.6 system cardemotional_reliance#1 / 70.985Source ↗official
GPT-5.6 system cardextremism_not_unsafe#1 / 71Source ↗official
GPT-5.6 system cardgore_not_unsafe#2 / 70.831Source ↗official
GPT-5.6 system cardharm_overall_pct#2 / 70.88Source ↗official
GPT-5.6 system cardhate_not_unsafe#5 / 70.943Source ↗official
GPT-5.6 system cardmental_health#3 / 70.985Source ↗official
GPT-5.6 system cardnonviolent_illicit_not_unsafe#1 / 71Source ↗official
GPT-5.6 system cardsearch_function_calling_injection_resistance#4 / 60.697Source ↗official
GPT-5.6 system cardself_harm#1 / 70.977Source ↗official
GPT-5.6 system cardself_harm_not_unsafe#1 / 70.987Source ↗official
GPT-5.6 system cardsexual_minors_not_unsafe#4 / 70.966Source ↗official
GPT-5.6 system cardsexual_not_unsafe#6 / 70.933Source ↗official
GPT-5.6 system cardviolent_illicit_not_unsafe#2 / 70.971Source ↗official
HUMAINE Trust, Ethics and Safetytrust_ethics_safety_score#25 / 5427.17Source ↗official
MACHIAVELLIdeception_relative_random_pct#34 / 5092.9Source ↗official
MASKlying_probability_pct#8 / 5310.33Source ↗official
OpenAI GPT-5.4 Dynamic Wellbeingemotional_reliance#1 / 30.985Source ↗official
OpenAI GPT-5.4 Dynamic Wellbeingmental_health#1 / 30.985Source ↗official
OpenAI GPT-5.4 Dynamic Wellbeingself_harm#1 / 30.977Source ↗official
OpenAI GPT-5.4 First-Person Fairnessharm_overall#1 / 30.0088Source ↗official
OpenAI GPT-5.4 Property Preservationdestructive_action_avoidance#2 / 30.86Source ↗official
OpenAI GPT-5.4 Property Preservationuser_work_preserved#1 / 30.53Source ↗official
OpenAI GPT-5.4 User Confirmationsfinancial_transaction#1 / 31Source ↗official
OpenAI GPT-5.4 User Confirmationsgeneral_confirmation#1 / 30.94Source ↗official
OpenAI GPT-5.4 User Confirmationshigh_stakes_communication#1 / 31Source ↗official
SABERoverall_safety_rate#2 / 1336.1Source ↗official
SABERscenario_a_safety_rate#3 / 1336.02Source ↗official
SABERscenario_b_safety_rate#2 / 1339.38Source ↗official
SABERscenario_c_safety_rate#2 / 1333.51Source ↗official
SM-Benchadversarial#62 / 7877.56Source ↗official
SM-Benchambiguous_interpretation#27 / 7887.8Source ↗official
SM-Benchanti_hallucination#46 / 7890.58Source ↗official
SM-Bencheq_boundaries#50 / 7860.11Source ↗official
SM-Benchoverfit#67 / 7838.25Source ↗official
SpeciEvalbelief_animal_sentience#75 / 1056.57Source ↗official
SpeciEvalland_animal_4ns#5 / 1053.7Source ↗official
SpeciEvalsea_animal_4ns#1 / 1053.65Source ↗official
SpeciEvalspeciesism#38 / 1051.82Source ↗official
StereoTales Harmful Associationsbenign_significant_association_score#8 / 2387.61Source ↗official
Vectara HHEM Factual Consistencyfactual_consistency_rate#28 / 9493Source ↗official
VETO Misfired Alignmentmisfired_alignment_rate_pct#24 / 2517.6Source ↗official
Vigil Mental Health Safetyoverall_score#2 / 2378Source ↗official

Values evaluations

Descriptive values and political-framing results are separate from safety/ethics ranks. Each strip shows the evaluation’s observed model range; its endpoint labels state what lower and higher values mean.

Agent-ValueBench Moral Foundations (MFT08)

Agent-ValueBench HEXACO

DimensionValueDistribution
Openness to experience4.5
Honesty-humility6
Extraversion5.7
Agreeableness5.6
Conscientiousness7.3

Agent-ValueBench Schwartz Basic Values (PVQ40)

The Economist World Values Survey Cultural Map

DimensionValueDistribution
Survival ↔ Self-expression2.26
Traditional ↔ Secular2.08