← Models

Model profile

Nemotron 3 Super 120B A12B

NVIDIAdeveloper
2026-03-10release date
#168 / 333Safety rank
#519 / 645Freedom rank
2discovery sources

Evidence summary

Safety. Nemotron 3 Super 120B A12B has an estimated Safety rank of #168; its 90% source-sensitivity interval is #90–#264. Its behavior-only rank is #175; company governance moves the combined estimate to #168. Published Safety evidence spans 11 eval lineages and 6 of 7 components. Its strongest relative result is Enkrypt AI Safety Leaderboard (harmful_attack_non_success_rate, #45 of 248); its weakest is Pokee-Isaac model card — DTAP (benign_task_success_rate, #6 of 6).

Freedom. Nemotron 3 Super 120B A12B has an estimated Freedom rank of #519; its 90% source-sensitivity interval is #325–#614. Published Freedom evidence spans 5 eval lineages and 1 of 1 components. Its strongest relative result is Enkrypt AI Safety Leaderboard (cbrn_attack_non_success_rate, #8 of 248); its weakest is SM-Bench (overfit, #81 of 84).

Compare this model

Only models sharing at least one published sub-eval are listed.

Official and reference links

Safety evals

Rank and direction are specific to the Safety portfolio. Black marks the observed result; the grey dot marks the value implied by the Safety rank. Values stay on each source’s native scale.

Eval / sub-evalRankValueDistributionSource
AA-Omnisciencehallucination_rate#234 / 3450.8696Source ↗official
Arena Factuality — Text Arena (factuality-only weighting)factuality_bt_rating#99 / 1111397.0Source ↗official
BullshitBench v2clear_pushback_rate#41 / 1170.4375Source ↗official
Cisco AI Defense Rolling Single-Turn Leaderboardsingle_turn_attack_success_rate#56 / 10440.43Source ↗official
DystopiaBenchbasaglia_score#29 / 5065.33Source ↗official
DystopiaBenchbaudrillard_score#22 / 5053.23Source ↗official
DystopiaBenchhuxley_score#23 / 5069.47Source ↗official
DystopiaBenchlaguardia_score#26 / 5067.17Source ↗official
DystopiaBenchorwell_score#25 / 5068.63Source ↗official
DystopiaBenchpetrov_score#34 / 5075.17Source ↗official
Enkrypt AI Safety Leaderboardbias_attack_non_success_rate#141 / 24814.73Source ↗official
Enkrypt AI Safety Leaderboardcbrn_attack_non_success_rate#241 / 24850.83Source ↗official
Enkrypt AI Safety Leaderboardharmful_attack_non_success_rate#45 / 24892.22Source ↗official
Enkrypt AI Safety Leaderboardtoxicity_attack_non_success_rate#62 / 24698.18Source ↗official
MT-JailBench CrescendoXsafety_score#5 / 2144.03Source ↗official
Pokee-Isaac model card — DTAPbenign_task_success_rate#6 / 60.633Source ↗official
Pokee-Isaac model card — DTAPcombined_attack_success_rate#5 / 60.604Source ↗official
RefusalBenchyouden_j#12 / 190.06383Source ↗official
SM-Benchadversarial#79 / 8471.22Source ↗official
SM-Benchambiguous_interpretation#80 / 8463.39Source ↗official
SM-Benchanti_hallucination#52 / 8490.58Source ↗official
SM-Bencheq_boundaries#75 / 8452.53Source ↗official
SM-Benchoverfit#81 / 8421.86Source ↗official
TACbase_welfare_rate#32 / 8729.5Source ↗self run

Freedom evals

Rank and direction are specific to the Freedom portfolio. Black marks the observed result; the grey dot marks the value implied by the Freedom rank. Values stay on each source’s native scale.

Eval / sub-evalRankValueDistributionSource
Cisco AI Defense Rolling Single-Turn Leaderboardsingle_turn_attack_success_rate#49 / 10440.43Source ↗official
Enkrypt AI Safety Leaderboardcbrn_attack_non_success_rate#8 / 24850.83Source ↗official
Enkrypt AI Safety Leaderboardharmful_attack_non_success_rate#202 / 24892.22Source ↗official
Enkrypt AI Safety Leaderboardtoxicity_attack_non_success_rate#180 / 24698.18Source ↗official
MT-JailBench CrescendoXsafety_score#17 / 2144.03Source ↗official
SM-Benchadversarial#6 / 8471.22Source ↗official
SM-Bencheq_boundaries#75 / 8452.53Source ↗official
SM-Benchoverfit#81 / 8421.86Source ↗official
SpeechMap model completioncomplete_pct#136 / 18142.5Source ↗official