← Models

Model profile

OLMo 2 13B Instruct

Ai2developer
2024-11-26release date
#290 / 333Safety rank
#180 / 645Freedom rank

Evidence summary

Safety. OLMo 2 13B Instruct has an estimated Safety rank of #290; its 90% source-sensitivity interval is #31–#321. Its behavior-only rank is #298; company governance moves the combined estimate to #290. Published Safety evidence spans 6 eval lineages and 5 of 7 components. Its strongest relative result is AILuminate General Purpose AI Chat (sex_related_crimes, #7 of 32); its weakest is HELM Safety (anthropic_red_team, #76 of 80).

Freedom. OLMo 2 13B Instruct has an estimated Freedom rank of #180; its 90% source-sensitivity interval is #43–#553. Published Freedom evidence spans 5 eval lineages and 1 of 1 components. Its strongest relative result is HELM Safety (anthropic_red_team, #5 of 80); its weakest is AILuminate General Purpose AI Chat (sex_related_crimes, #26 of 32).

Compare this model

Only models sharing at least one published sub-eval are listed.

Official and reference links

Safety evals

Rank and direction are specific to the Safety portfolio. Black marks the observed result; the grey dot marks the value implied by the Safety rank. Values stay on each source’s native scale.

Eval / sub-evalRankValueDistributionSource
AILuminate General Purpose AI Chatchild_sexual_exploitation#12 / 3214.2Source ↗official
AILuminate General Purpose AI Chatdefamation#11 / 3211Source ↗official
AILuminate General Purpose AI Chathate#19 / 3210.7Source ↗official
AILuminate General Purpose AI Chatindiscriminate_weapons_cbrne#18 / 3225Source ↗official
AILuminate General Purpose AI Chatintellectual_property#11 / 328.2Source ↗official
AILuminate General Purpose AI Chatnon_violent_crimes#10 / 329.3Source ↗official
AILuminate General Purpose AI Chatprivacy#11 / 328.6Source ↗official
AILuminate General Purpose AI Chatsex_related_crimes#7 / 327.5Source ↗official
AILuminate General Purpose AI Chatsexual_content#11 / 3217.5Source ↗official
AILuminate General Purpose AI Chatspecialized_advice#11 / 3213.2Source ↗official
AILuminate General Purpose AI Chatsuicide_self_harm#9 / 328.5Source ↗official
AILuminate General Purpose AI Chatviolent_crimes#9 / 3212Source ↗official
HELM Safetyanthropic_red_team#76 / 800.907Source ↗official
HELM Safetybbq#75 / 800.704Source ↗official
HELM Safetyharmbench#69 / 800.519Source ↗official
HELM Safetysimple_safety_tests#76 / 800.81Source ↗official
HELM Safetyxstest#64 / 800.926Source ↗official

Freedom evals

Rank and direction are specific to the Freedom portfolio. Black marks the observed result; the grey dot marks the value implied by the Freedom rank. Values stay on each source’s native scale.

Eval / sub-evalRankValueDistributionSource
AILuminate General Purpose AI Chatchild_sexual_exploitation#21 / 3214.2Source ↗official
AILuminate General Purpose AI Chatdefamation#22 / 3211Source ↗official
AILuminate General Purpose AI Chathate#14 / 3210.7Source ↗official
AILuminate General Purpose AI Chatindiscriminate_weapons_cbrne#15 / 3225Source ↗official
AILuminate General Purpose AI Chatintellectual_property#21 / 328.2Source ↗official
AILuminate General Purpose AI Chatnon_violent_crimes#23 / 329.3Source ↗official
AILuminate General Purpose AI Chatprivacy#22 / 328.6Source ↗official
AILuminate General Purpose AI Chatsex_related_crimes#26 / 327.5Source ↗official
AILuminate General Purpose AI Chatsexual_content#22 / 3217.5Source ↗official
AILuminate General Purpose AI Chatspecialized_advice#22 / 3213.2Source ↗official
AILuminate General Purpose AI Chatsuicide_self_harm#24 / 328.5Source ↗official
AILuminate General Purpose AI Chatviolent_crimes#23 / 3212Source ↗official
HELM Safetyanthropic_red_team#5 / 800.907Source ↗official
HELM Safetyharmbench#12 / 800.519Source ↗official
HELM Safetysimple_safety_tests#5 / 800.81Source ↗official
HELM Safetyxstest#64 / 800.926Source ↗official