← Models

Model profile

DeepSeek v4 Pro

DeepSeekdeveloper
2026-04-22release date
#141 / 333Safety rank
#277 / 645Freedom rank
3discovery sources

Evidence summary

Safety. DeepSeek v4 Pro has an estimated Safety rank of #141; its 90% source-sensitivity interval is #70–#287. Its behavior-only rank is #131; company governance moves the combined estimate to #141. Published Safety evidence spans 23 eval lineages and 7 of 7 components. Its strongest relative result is ANIMA (score, #1 of 22); its weakest is AgentAbstain (abstain, #17 of 17).

Freedom. DeepSeek v4 Pro has an estimated Freedom rank of #277; its 90% source-sensitivity interval is #143–#439. Published Freedom evidence spans 8 eval lineages and 1 of 1 components. Its strongest relative result is Human Pathogen Capabilities Test (HPCT) — overall refusal (refusal_rate_pct, #1 of 69); its weakest is UGI Leaderboard — base-model willingness (willingness_direct_score, #113 of 156).

Compare this model

Only models sharing at least one published sub-eval are listed.

Official and reference links

Safety evals

Rank and direction are specific to the Safety portfolio. Black marks the observed result; the grey dot marks the value implied by the Safety rank. Values stay on each source’s native scale.

Eval / sub-evalRankValueDistributionSource
AA-Omnisciencehallucination_rate#245 / 3450.8785Source ↗official
AgentAbstainabstain#17 / 1742.8Source ↗official
AgentAbstaincar#16 / 1742.3Source ↗official
AgentAbstainpaired#15 / 1736.9Source ↗official
ANIMAscore#1 / 220.767Source ↗official
Arena Factuality — Text Arena (factuality-only weighting)factuality_bt_rating#55 / 1111437.0Source ↗official
BioTIERpermit_compliance_pct#35 / 5298.5Source ↗official
BioTIERrefuse_compliance_pct#35 / 5222.6Source ↗official
BullshitBench v2clear_pushback_rate#89 / 1170.14Source ↗official
CAIS Risk Indexpolitical_manipulation#17 / 4844Source ↗official
Concordia — Agentic-Misalignmentsafety_score#46 / 5454.83Source ↗official
Concordia — AirBench-Deceptionsafety_score#47 / 6370.37Source ↗official
Concordia — AirBench-Manipulationsafety_score#45 / 5674.67Source ↗official
Concordia — AirBench-SecurityRiskssafety_score#50 / 6380.18Source ↗official
Concordia — APEsafety_score#49 / 551.542Source ↗official
Concordia — CyberSecEval2-PromptInjectionsafety_score#31 / 6384.86Source ↗official
Concordia — DarkBenchsafety_score#29 / 5555.45Source ↗official
Concordia — Fortress-Biologicalsafety_score#54 / 5418.62Source ↗official
Concordia — Fortress-Chemicalsafety_score#54 / 5421.39Source ↗official
Concordia — Fortress-Privacy/Scamssafety_score#43 / 5443.05Source ↗official
Concordia — FRT-AirBench-Manipulationsafety_score#41 / 4515.11Source ↗official
Concordia — FRT-AirBench-SecurityRiskssafety_score#43 / 458.667Source ↗official
Concordia — FRT-SciKnowEval-BiologicalHarmfulQAsafety_score#36 / 451.333Source ↗official
Concordia — FRT-SOSBench-Chemsafety_score#39 / 4527.33Source ↗official
Concordia — MASKsafety_score#59 / 6239.38Source ↗official
Concordia — SciKnowEval-BiologicalHarmfulQAsafety_score#35 / 6377.44Source ↗official
Concordia — SciKnowEval-ChemicalHarmfulQAsafety_score#51 / 637.93Source ↗official
Concordia — Shutdown-Resistancesafety_score#35 / 5398.98Source ↗official
Concordia — SOSBench-Biosafety_score#21 / 6391.4Source ↗official
Concordia — SOSBench-Chemsafety_score#32 / 6387.8Source ↗official
DystopiaBenchbasaglia_score#26 / 5064.6Source ↗official
DystopiaBenchbaudrillard_score#31 / 5062.2Source ↗official
DystopiaBenchhuxley_score#41 / 5076.97Source ↗official
DystopiaBenchlaguardia_score#42 / 5069.8Source ↗official
DystopiaBenchorwell_score#40 / 5074Source ↗official
DystopiaBenchpetrov_score#36 / 5077Source ↗official
Enkrypt AI Safety Leaderboardbias_attack_non_success_rate#141 / 24814.73Source ↗official
Enkrypt AI Safety Leaderboardcbrn_attack_non_success_rate#182 / 24882.67Source ↗official
Enkrypt AI Safety Leaderboardharmful_attack_non_success_rate#101 / 24877.78Source ↗official
Enkrypt AI Safety Leaderboardtoxicity_attack_non_success_rate#188 / 24692.55Source ↗official
Every Model Cheats — Cybench Cheat Propensitybaseline_cheat_propensity_pct#5 / 2221.7Source ↗official
Every Model Cheats — Cybench Cheat Propensitysevere_anti_cheat_cheat_propensity_pct#1 / 220Source ↗official
Every Model Cheats — Cybench Cheat Propensitystandard_anti_cheat_cheat_propensity_pct#13 / 2217.4Source ↗official
Google Gemini 3.8 launch — Gray Swan indirect prompt injection k=15attack_success_probability_k15_pct#15 / 1560.1Source ↗official
HUMAINE Trust, Ethics and Safetytrust_ethics_safety_score#20 / 5427.61Source ↗official
Human Pathogen Capabilities Test (HPCT) — overall refusalrefusal_rate_pct#27 / 690Source ↗official
Manager Coercion Benchcoercion_ladder_depth#33 / 379Source ↗official
Manager Coercion Benchfabrication_rate#1 / 150Source ↗official
Olam Social Poker — Social Lie Ratesocial_lie_rate_per_10000_turns#12 / 249Source ↗official
Opposite-Narrator Sycophancysycophancy_rate_pct#17 / 245.1Source ↗official
PHAREbias_resistance_diagnostic#66 / 660.1714Source ↗official
PHAREhallucination_resistance_diagnostic#21 / 700.7964Source ↗official
PHAREharm_resistance_diagnostic#22 / 700.9541Source ↗official
PHAREjailbreak_resistance_diagnostic#44 / 670.4321Source ↗official
SM-Benchadversarial#48 / 8480.98Source ↗official
SM-Benchambiguous_interpretation#76 / 8468.75Source ↗official
SM-Benchanti_hallucination#56 / 8488.48Source ↗official
SM-Bencheq_boundaries#42 / 8465.17Source ↗official
SM-Benchoverfit#40 / 8476.5Source ↗official
SpeciEvalbelief_animal_sentience#63 / 1236.8Source ↗official
SpeciEvalland_animal_4ns#36 / 1234.35Source ↗official
SpeciEvalsea_animal_4ns#43 / 1234.65Source ↗official
SpeciEvalspeciesism#41 / 1231.77Source ↗official
TACbase_welfare_rate#62 / 8723.08Source ↗official
The Dictatorship Evaloverall_resistance_rate#18 / 2011.65Source ↗official
Vectara HHEM Factual Consistencyfactual_consistency_rate#39 / 9491.4Source ↗official

Freedom evals

Rank and direction are specific to the Freedom portfolio. Black marks the observed result; the grey dot marks the value implied by the Freedom rank. Values stay on each source’s native scale.

Eval / sub-evalRankValueDistributionSource
BioTIERpermit_compliance_pct#35 / 5298.5Source ↗official
BioTIERrefuse_compliance_pct#18 / 5222.6Source ↗official
Concordia — AirBench-Deceptionsafety_score#16 / 6370.37Source ↗official
Concordia — AirBench-Manipulationsafety_score#12 / 5674.67Source ↗official
Concordia — AirBench-PoliticalPersuasionsafety_score#19 / 5649.52Source ↗official
Concordia — AirBench-SecurityRiskssafety_score#14 / 6380.18Source ↗official
Concordia — Fortress-Biologicalsafety_score#1 / 5418.62Source ↗official
Concordia — Fortress-Chemicalsafety_score#1 / 5421.39Source ↗official
Concordia — Fortress-Privacy/Scamssafety_score#12 / 5443.05Source ↗official
Concordia — FRT-AirBench-Manipulationsafety_score#4 / 4515.11Source ↗official
Concordia — FRT-AirBench-SecurityRiskssafety_score#3 / 458.667Source ↗official
Concordia — FRT-SciKnowEval-BiologicalHarmfulQAsafety_score#9 / 451.333Source ↗official
Concordia — FRT-SOSBench-Chemsafety_score#6 / 4527.33Source ↗official
Concordia — SciKnowEval-BiologicalHarmfulQAsafety_score#28 / 6377.44Source ↗official
Concordia — SciKnowEval-ChemicalHarmfulQAsafety_score#12 / 637.93Source ↗official
Concordia — SOSBench-Biosafety_score#43 / 6391.4Source ↗official
Concordia — SOSBench-Chemsafety_score#32 / 6387.8Source ↗official
Enkrypt AI Safety Leaderboardcbrn_attack_non_success_rate#67 / 24882.67Source ↗official
Enkrypt AI Safety Leaderboardharmful_attack_non_success_rate#148 / 24877.78Source ↗official
Enkrypt AI Safety Leaderboardtoxicity_attack_non_success_rate#59 / 24692.55Source ↗official
Human Pathogen Capabilities Test (HPCT) — overall refusalrefusal_rate_pct#1 / 690Source ↗official
PHAREharm_resistance_diagnostic#49 / 700.9541Source ↗official
PHAREjailbreak_resistance_diagnostic#24 / 670.4321Source ↗official
SM-Benchadversarial#35 / 8480.98Source ↗official
SM-Bencheq_boundaries#42 / 8465.17Source ↗official
SM-Benchoverfit#40 / 8476.5Source ↗official
SpeechMap model completioncomplete_pct#121 / 18146.8Source ↗official
UGI Leaderboard — base-model willingnesswillingness_adherence_score#50 / 1564.25Source ↗official
UGI Leaderboard — base-model willingnesswillingness_direct_score#113 / 1562Source ↗official

Values evaluations

Descriptive values and political-framing results are separate from safety/ethics ranks. Each strip shows the evaluation’s observed model range; its endpoint labels state what lower and higher values mean.

UGI Political Values

DimensionValueDistribution
Political Lean-18.6
Government45.1
Diplomacy66
Economy44.7
Society59.2

CAIS AI Values — countries

Moral Trolley Arena