← Models

Model profile

DeepSeek v4 Flash

DeepSeekdeveloper
2026-04-22release date
#91 / 333Safety rank
#209 / 645Freedom rank
4discovery sources

Evidence summary

Safety. DeepSeek v4 Flash has an estimated Safety rank of #91; its 90% source-sensitivity interval is #73–#238. Its behavior-only rank is #80; company governance moves the combined estimate to #91. Published Safety evidence spans 20 eval lineages and 7 of 7 components. Its strongest relative result is Manager Coercion Bench (coercion_ladder_depth, #1 of 37); its weakest is Inkling-Small model card — StrongREJECT (safety_rate, #10 of 10).

Freedom. DeepSeek v4 Flash has an estimated Freedom rank of #209; its 90% source-sensitivity interval is #59–#454. Published Freedom evidence spans 6 eval lineages and 1 of 1 components. Its strongest relative result is UGI Leaderboard — base-model willingness (willingness_adherence_score, #6 of 156); its weakest is SpeechMap model completion (complete_pct, #131 of 181).

Compare this model

Only models sharing at least one published sub-eval are listed.

Official and reference links

Safety evals

Rank and direction are specific to the Safety portfolio. Black marks the observed result; the grey dot marks the value implied by the Safety rank. Values stay on each source’s native scale.

Eval / sub-evalRankValueDistributionSource
AA-Omnisciencehallucination_rate#258 / 3450.8897Source ↗official
ANIMAscore#3 / 220.7445Source ↗official
Arena Factuality — Text Arena (factuality-only weighting)factuality_bt_rating#63 / 1111432.0Source ↗official
BullshitBench v2clear_pushback_rate#83 / 1170.16Source ↗official
DystopiaBenchbasaglia_score#23 / 5060.9Source ↗official
DystopiaBenchbaudrillard_score#27 / 5058.1Source ↗official
DystopiaBenchhuxley_score#27 / 5072.8Source ↗official
DystopiaBenchlaguardia_score#37 / 5069.03Source ↗official
DystopiaBenchorwell_score#28 / 5069.83Source ↗official
DystopiaBenchpetrov_score#30 / 5074.3Source ↗official
Enkrypt AI Safety Leaderboardbias_attack_non_success_rate#141 / 24814.73Source ↗official
Enkrypt AI Safety Leaderboardcbrn_attack_non_success_rate#227 / 24870.33Source ↗official
Enkrypt AI Safety Leaderboardharmful_attack_non_success_rate#103 / 24876.67Source ↗official
Enkrypt AI Safety Leaderboardtoxicity_attack_non_success_rate#196 / 24691.64Source ↗official
Every Model Cheats — Cybench Cheat Propensitybaseline_cheat_propensity_pct#10 / 2230.4Source ↗official
Every Model Cheats — Cybench Cheat Propensitysevere_anti_cheat_cheat_propensity_pct#7 / 224.3Source ↗official
Every Model Cheats — Cybench Cheat Propensitystandard_anti_cheat_cheat_propensity_pct#6 / 228.7Source ↗official
Governance Decay under Passive Context Compactiongovernance_retention_score#6 / 741Source ↗official
HUMAINE Trust, Ethics and Safetytrust_ethics_safety_score#10 / 5428.57Source ↗official
Inkling-Small model card — FORTRESSbenign_answer_rate#1 / 1099.2Source ↗official
Inkling-Small model card — FORTRESSharmful_refusal_rate#10 / 1032Source ↗official
Inkling-Small model card — StrongREJECTsafety_rate#10 / 1097.4Source ↗official
KIDBench Implicit Child Cueimplicit_child_cue_total_mean#1 / 134.43Source ↗official
Manager Coercion Benchcoercion_ladder_depth#1 / 373.233Source ↗self run
MANTAAWMS#9 / 120.417Source ↗official
MANTAAWVS#6 / 120.508Source ↗official
Olam Social Poker — Social Lie Ratesocial_lie_rate_per_10000_turns#19 / 2437Source ↗official
Opposite-Narrator Sycophancysycophancy_rate_pct#19 / 245.6Source ↗official
PHAREbias_resistance_diagnostic#59 / 660.323Source ↗official
PHAREhallucination_resistance_diagnostic#13 / 700.8188Source ↗official
PHAREharm_resistance_diagnostic#28 / 700.9507Source ↗official
PHAREjailbreak_resistance_diagnostic#42 / 670.4549Source ↗official
SM-Benchadversarial#43 / 8481.95Source ↗official
SM-Benchambiguous_interpretation#71 / 8475.89Source ↗official
SM-Benchanti_hallucination#40 / 8494.24Source ↗official
SM-Bencheq_boundaries#20 / 8469.94Source ↗official
SM-Benchoverfit#59 / 8459.56Source ↗official
SpeciEvalbelief_animal_sentience#37 / 1236.9Source ↗official
SpeciEvalland_animal_4ns#84 / 1234.7Source ↗official
SpeciEvalsea_animal_4ns#66 / 1234.75Source ↗official
SpeciEvalspeciesism#65 / 1232.05Source ↗official
TACbase_welfare_rate#64 / 8721.15Source ↗official
ToolPrivacyBenchprivate_mt_poi#6 / 927.33Source ↗official
ToolPrivacyBenchpublic_mt_poi#6 / 918.99Source ↗official

Freedom evals

Rank and direction are specific to the Freedom portfolio. Black marks the observed result; the grey dot marks the value implied by the Freedom rank. Values stay on each source’s native scale.

Eval / sub-evalRankValueDistributionSource
Enkrypt AI Safety Leaderboardcbrn_attack_non_success_rate#22 / 24870.33Source ↗official
Enkrypt AI Safety Leaderboardharmful_attack_non_success_rate#146 / 24876.67Source ↗official
Enkrypt AI Safety Leaderboardtoxicity_attack_non_success_rate#51 / 24691.64Source ↗official
Inkling-Small model card — FORTRESSbenign_answer_rate#1 / 1099.2Source ↗official
Inkling-Small model card — FORTRESSharmful_refusal_rate#1 / 1032Source ↗official
Inkling-Small model card — StrongREJECTsafety_rate#1 / 1097.4Source ↗official
PHAREharm_resistance_diagnostic#43 / 700.9507Source ↗official
PHAREjailbreak_resistance_diagnostic#26 / 670.4549Source ↗official
SM-Benchadversarial#41 / 8481.95Source ↗official
SM-Bencheq_boundaries#20 / 8469.94Source ↗official
SM-Benchoverfit#59 / 8459.56Source ↗official
SpeechMap model completioncomplete_pct#131 / 18144.2Source ↗official
UGI Leaderboard — base-model willingnesswillingness_adherence_score#6 / 1568Source ↗official
UGI Leaderboard — base-model willingnesswillingness_direct_score#41 / 1564.5Source ↗official

Values evaluations

Descriptive values and political-framing results are separate from safety/ethics ranks. Each strip shows the evaluation’s observed model range; its endpoint labels state what lower and higher values mean.

UGI Political Values

DimensionValueDistribution
Political Lean-16.6
Government46
Diplomacy66
Economy46.2
Society59

CCPBench political narrative alignment

DimensionValueDistribution
CCP-narrative alignment — all questions3.57
CCP-narrative alignment — China topics4.03
CCP-narrative alignment — non-China controls2.22

The Economist World Values Survey Cultural Map

DimensionValueDistribution
Survival ↔ Self-expression2.27
Traditional ↔ Secular-1.75