← Models

Model profile

Korani 13B

1eval lineages

Evidence summary

Published evidence spans 1 evals and 1 of 7 behavior components. Its strongest relative result is HyperCLOVA X Toxic Continuation Panels (kold_toxic_count, #4 of 7); its weakest is HyperCLOVA X Toxic Continuation Panels (realt-toxicprompts_toxicity, #6 of 7).

Compare this model

Only models sharing at least one published sub-eval are listed.

Published eval results

Rank is within that sub-eval. Black marks the observed result; the grey dot marks the value implied by the global rank. Values stay on each source’s native scale.

Eval / sub-evalRankValueDistributionBetterSource
HyperCLOVA X Toxic Continuation Panelskold_toxic_count#4 / 70.008↓ lowerSource ↗official
HyperCLOVA X Toxic Continuation Panelskold_toxicity#5 / 70.1329↓ lowerSource ↗official
HyperCLOVA X Toxic Continuation Panelsrealt-toxicprompts_toxic_count#6 / 70.026↓ lowerSource ↗official
HyperCLOVA X Toxic Continuation Panelsrealt-toxicprompts_toxicity#6 / 70.1076↓ lowerSource ↗official