← Models

Model profile

Marin 8B Instruct

Marindeveloper
2025-05-14release date
#254 / 267overall rank
5eval lineages

Evidence summary

Marin 8B Instruct has an estimated overall rank of #254; its 90% source-sensitivity interval is #158–#258. Its behavior-only rank is #258; company governance moves the combined estimate to #254. Published evidence spans 5 evals and 3 of 7 behavior components. Its strongest relative result is HELM Safety (harmbench, #39 of 80); its weakest is HELM Safety (bbq, #71 of 80).

Compare this model

Only models sharing at least one published sub-eval are listed.

Official and reference links

Published eval results

Rank is within that sub-eval. Black marks the observed result; the grey dot marks the value implied by the global rank. Values stay on each source’s native scale.

Eval / sub-evalRankValueDistributionBetterSource
HELM Safetyanthropic_red_team#63 / 800.97↑ higherSource ↗official
HELM Safetybbq#71 / 800.755↑ higherSource ↗official
HELM Safetyharmbench#39 / 800.751↑ higherSource ↗official
HELM Safetysimple_safety_tests#66 / 800.958↑ higherSource ↗official
HELM Safetyxstest#67 / 800.917↑ higherSource ↗official