← Models

Model profile

GPT Oss 20B

OpenAIdeveloper
2025-08-05release date
#81 / 267overall rank
16eval lineages
2discovery sources

Evidence summary

GPT Oss 20B has an estimated overall rank of #81; its 90% source-sensitivity interval is #39–#180. Its behavior-only rank is #88; company governance moves the combined estimate to #81. Published evidence spans 16 evals and 7 of 7 behavior components. Its strongest relative result is HELM Safety (simple_safety_tests, #1 of 80); its weakest is Enkrypt AI Safety Leaderboard (toxicity_attack_non_success_rate, #239 of 258).

Compare this model

Only models sharing at least one published sub-eval are listed.

Official and reference links

Published eval results

Rank is within that sub-eval. Black marks the observed result; the grey dot marks the value implied by the global rank. Values stay on each source’s native scale.

Eval / sub-evalRankValueDistributionBetterSource
AA-Omnisciencehallucination_rate#212 / 3110.866↓ lowerSource ↗official
AIRBench 2024 Safety Scenariossafety_scenarios#10 / 800.86↑ higherSource ↗official
Cisco AI Defense Rolling Single-Turn Leaderboardsingle_turn_attack_success_rate#30 / 10514.17↓ lowerSource ↗official
DystopiaBenchbasaglia_score#42 / 5070.13↓ lowerSource ↗official
DystopiaBenchbaudrillard_score#37 / 5066.3↓ lowerSource ↗official
DystopiaBenchhuxley_score#37 / 5075.47↓ lowerSource ↗official
DystopiaBenchlaguardia_score#31 / 5068.27↓ lowerSource ↗official
DystopiaBenchorwell_score#35 / 5073.53↓ lowerSource ↗official
DystopiaBenchpetrov_score#28 / 5073.53↓ lowerSource ↗official
Enkrypt AI Safety Leaderboardbias_attack_non_success_rate#54 / 26027.65↑ higherSource ↗official
Enkrypt AI Safety Leaderboardcbrn_attack_non_success_rate#213 / 26081↑ higherSource ↗official
Enkrypt AI Safety Leaderboardharmful_attack_non_success_rate#94 / 26080.56↑ higherSource ↗official
Enkrypt AI Safety Leaderboardtoxicity_attack_non_success_rate#239 / 25880↑ higherSource ↗official
FORTRESSaverage_risk_score#15 / 4917.62↓ lowerSource ↗official
FORTRESSover_refusal_score#41 / 4611.94↓ lowerSource ↗official
HELM Safetyanthropic_red_team#10 / 800.997↑ higherSource ↗official
HELM Safetybbq#13 / 800.967↑ higherSource ↗official
HELM Safetyharmbench#2 / 800.987↑ higherSource ↗official
HELM Safetysimple_safety_tests#1 / 801↑ higherSource ↗official
HELM Safetyxstest#72 / 800.872↑ higherSource ↗official
Manager Coercion Benchcoercion_ladder_depth#11 / 317.9↓ lowerSource ↗self run
MASKlying_probability_pct#10 / 5313.54↓ lowerSource ↗official
ODCV-Benchaverage_severity#7 / 121.475↓ lowerSource ↗official
ODCV-Benchmisalignment_rate#5 / 1228.75↓ lowerSource ↗official
SpeciEvalbelief_animal_sentience#58 / 1026.75↑ higherSource ↗official
SpeciEvalland_animal_4ns#12 / 1024.2↓ lowerSource ↗official
SpeciEvalsea_animal_4ns#59 / 1024.83↓ lowerSource ↗official
SpeciEvalspeciesism#40 / 1021.9↓ lowerSource ↗official
Vigil Mental Health Safetyoverall_score#17 / 2331↑ higherSource ↗official

Values evaluations

Descriptive values and political-framing results are separate from safety/ethics ranks. Each strip shows the evaluation’s observed model range; its endpoint labels state what lower and higher values mean.

UGI Political Values

DimensionValueDistribution
Political Lean-16.9
Government47.4
Diplomacy64.7
Economy45.2
Society60.1