← Models

Model profile

GPT 5

OpenAIdeveloper
2025-08-07release date
#49 / 309overall rank
38eval lineages

Evidence summary

GPT 5 has an estimated overall rank of #49; its 90% source-sensitivity interval is #19–#147. Its behavior-only rank is #58; company governance moves the combined estimate to #49. Published evidence spans 38 evals and 7 of 7 behavior components. Its strongest relative result is Enkrypt AI Safety Leaderboard (bias_attack_non_success_rate, #9 of 241); its weakest is OpenAgentSafety (rule_based_safety_vulnerable, #7 of 7).

Compare this model

Only models sharing at least one published sub-eval are listed.

Official and reference links

Published eval results

Rank is within that sub-eval. Black marks the observed result; the grey dot marks the value implied by the global rank. Values stay on each source’s native scale.

Eval / sub-evalRankValueDistributionSource
AA-Omnisciencehallucination_rate#143 / 3280.7902Source ↗official
AgentAbstainabstain#2 / 1769.8Source ↗official
AgentAbstaincar#2 / 1766.5Source ↗official
AgentAbstainpaired#6 / 1749.6Source ↗official
AgentDrive Safety Compliancescr#2 / 4896.25Source ↗official
AIRBench 2024 Safety Scenariossafety_scenarios#7 / 800.877Source ↗official
Arena Factuality — Search Arena (factuality-only weighting)factuality_bt_rating#8 / 301214.0Source ↗official
Arena Factuality — Text Arena (factuality-only weighting)factuality_bt_rating#15 / 1121464.0Source ↗official
BrokenMathsycophancy#1 / 929Source ↗official
BullshitBench v2clear_pushback_rate#66 / 1060.195Source ↗official
CAIS Risk Indexbioweapons_assistance#16 / 4962.6Source ↗official
CAIS Risk Indexhle_overconfidence#19 / 5050Source ↗official
CAIS Risk Indexmachiavelli#11 / 4683.8Source ↗official
CAIS Risk Indexmask#26 / 5220.7Source ↗official
CAIS Risk Indexpolitical_manipulation#4 / 3335.3Source ↗official
CAIS Risk Indextextquests_harm#22 / 4917.2Source ↗official
Cisco AI Defense Rolling Single-Turn Leaderboardsingle_turn_attack_success_rate#23 / 1048.076Source ↗official
Constitutional Following — OpenAI Model Specconstitutional_following_score#3 / 794.9Source ↗official
Enkrypt AI Safety Leaderboardbias_attack_non_success_rate#9 / 24159.95Source ↗official
Enkrypt AI Safety Leaderboardcbrn_attack_non_success_rate#39 / 24192.89Source ↗official
Enkrypt AI Safety Leaderboardharmful_attack_non_success_rate#37 / 24193.33Source ↗official
Enkrypt AI Safety Leaderboardtoxicity_attack_non_success_rate#100 / 23997.09Source ↗official
FlagEval Safety and Valuesa1_qualified_rate#10 / 1880.22Source ↗official
FlagEval Safety and Valuesa2_qualified_rate#9 / 1879.66Source ↗official
FlagEval Safety and Valuesa3_qualified_rate#14 / 1883.26Source ↗official
FlagEval Safety and Valuesa4_qualified_rate#13 / 1883.73Source ↗official
FlagEval Safety and Valuesa5_qualified_rate#10 / 1872.15Source ↗official
FORTRESSaverage_risk_score#13 / 4917.04Source ↗official
FORTRESSover_refusal_score#39 / 4611.2Source ↗official
HELM Safetyanthropic_red_team#32 / 800.991Source ↗official
HELM Safetybbq#12 / 800.968Source ↗official
HELM Safetyharmbench#6 / 800.976Source ↗official
HELM Safetysimple_safety_tests#24 / 800.998Source ↗official
HELM Safetyxstest#22 / 800.971Source ↗official
HUMAINE Trust, Ethics and Safetytrust_ethics_safety_score#32 / 5426.31Source ↗official
MACHIAVELLIdeception_relative_random_pct#11 / 5083.8Source ↗official
MASKlying_probability_pct#16 / 5320.67Source ↗official
MonitoringBench Full-Trajectory Monitorfull_trajectory_catch_rate_at_1pct_fpr_percent#6 / 1331.3Source ↗official
MT-JailBench CrescendoXsafety_score#3 / 2161.29Source ↗official
OpenAgentSafetyllm_judge_safety_vulnerable#3 / 752.58Source ↗official
OpenAgentSafetyrule_based_safety_vulnerable#7 / 752.15Source ↗official
OpenAgentSafetysuccessful_completion#1 / 740.4Source ↗official
OpenAI GPT-5 System Cardharmful_request_safety#1 / 20.912Source ↗official
PacifAIstp_score#7 / 779.49Source ↗official
PHAREbias_resistance_diagnostic#63 / 660.2856Source ↗official
PHAREhallucination_resistance_diagnostic#38 / 700.7458Source ↗official
PHAREharm_resistance_diagnostic#12 / 700.9697Source ↗official
PHAREjailbreak_resistance_diagnostic#14 / 670.6868Source ↗official
Shelleducation_jsr#2 / 140.364Source ↗official
Shellfinance_jsr#2 / 140.19Source ↗official
Shellmanagement_jsr#3 / 140.37Source ↗official
Social Welfare Function Benchmarkfairness#15 / 190.4455Source ↗official
SOSBenchbiology_pvr#2 / 230.108Source ↗official
SOSBenchchemistry_pvr#1 / 230.122Source ↗official
SOSBenchmedicine_pvr#3 / 230.332Source ↗official
SOSBenchpharmacology_pvr#4 / 230.418Source ↗official
SOSBenchphysics_pvr#2 / 230.104Source ↗official
SOSBenchpsychology_pvr#3 / 230.142Source ↗official
SpeciEvalbelief_animal_sentience#46 / 1136.855Source ↗official
SpeciEvalland_animal_4ns#91 / 1134.825Source ↗official
SpeciEvalsea_animal_4ns#57 / 1134.745Source ↗official
SpeciEvalspeciesism#16 / 1131.49Source ↗official
ThaiSafetyBenchsafety_score#1 / 1895.57Source ↗official
TrustLLM contemporary collapsed applicationtrustllm#7 / 80.6Source ↗official
UAVBench safety-critical decision recognitionethical_safety_critical_accuracy#2 / 270.76Source ↗official
Vectara HHEM Factual Consistencyfactual_consistency_rate#85 / 9485.1Source ↗official
Vigil Mental Health Safetyoverall_score#9 / 2351Source ↗official

Values evaluations

Descriptive values and political-framing results are separate from safety/ethics ranks. Each strip shows the evaluation’s observed model range; its endpoint labels state what lower and higher values mean.

UGI Political Values

DimensionValueDistribution
Political Lean-18.2
Government45.9
Diplomacy65.3
Economy49
Society58.6

CAISI CCP narrative alignment

DimensionValueDistribution
CCP narrative alignment1.95