← Evals

Evaluation profile

AnimalHarmBench

1sub-evals
3.4%Safety weight
0%Freedom weight
1components

Weights below are portfolio-specific global index weights.

Model score (higher is better)Predicted score

About this eval

Whether model-generated text increases or decreases risk of animal harm.

Included in the behavior ranking.

Sub-evals

MeasureComponentDirection by rankingSafety weightFreedom weight
scoreahb.csv:scoreMeasures whether the model recognizes and avoids harms to non-human animals across animal-welfare scenarios.Safety: nonhuman_ethics:1.000ahb.csvSafety: higher3.4%

score

Measures whether the model recognizes and avoids harms to non-human animals across animal-welfare scenarios.

RankModelValueRelative performanceProvenance
1mistral-large0.068official
2gemini-1.5-pro0.066official
3gemini-1.5-flash0.05official
4claude-3-opus0.043official
5deepseek-v30.04official
6claude-3.5-haiku0.02official
7claude-3.5-sonnet0.018official
8gpt-4o0.011official
9gpt-4o-mini0.002official
10llama-3.3-70b-instruct-0.015official