← Models

Model profile

Granite 4.0 Micro

IBMdeveloper
2025-10-02release date
#286 / 309overall rank
9eval lineages

Evidence summary

Granite 4.0 Micro has an estimated overall rank of #286; its 90% source-sensitivity interval is #176–#297. Its behavior-only rank is #292; company governance moves the combined estimate to #286. Published evidence spans 9 evals and 6 of 7 behavior components. Its strongest relative result is Enkrypt AI Safety Leaderboard (toxicity_attack_non_success_rate, #25 of 239); its weakest is AA-Omniscience (hallucination_rate, #315 of 328).

Compare this model

Only models sharing at least one published sub-eval are listed.

Official and reference links

Published eval results

Rank is within that sub-eval. Black marks the observed result; the grey dot marks the value implied by the global rank. Values stay on each source’s native scale.

Eval / sub-evalRankValueDistributionSource
AA-Omnisciencehallucination_rate#315 / 3280.9577Source ↗official
AIRBench 2024 Safety Scenariossafety_scenarios#42 / 800.661Source ↗official
Enkrypt AI Safety Leaderboardbias_attack_non_success_rate#129 / 24114.99Source ↗official
Enkrypt AI Safety Leaderboardcbrn_attack_non_success_rate#84 / 24190.17Source ↗official
Enkrypt AI Safety Leaderboardharmful_attack_non_success_rate#51 / 24188.89Source ↗official
Enkrypt AI Safety Leaderboardtoxicity_attack_non_success_rate#25 / 23999.45Source ↗official
HELM Safetyanthropic_red_team#44 / 800.987Source ↗official
HELM Safetybbq#69 / 800.779Source ↗official
HELM Safetyharmbench#47 / 800.695Source ↗official
HELM Safetysimple_safety_tests#68 / 800.945Source ↗official
HELM Safetyxstest#73 / 800.869Source ↗official
UAVBench safety-critical decision recognitionethical_safety_critical_accuracy#25 / 270.54Source ↗official