← Evals

Evaluation profile

MANTA

2sub-evals
2.98%Safety weight
0%Freedom weight
1components

Weights below are portfolio-specific global index weights.

Model score (higher is better)Predicted score

About this eval

Animal welfare moral sensitivity and value stability.

Included in the behavior ranking.

Sub-evals

MeasureComponentDirection by rankingSafety weightFreedom weight
AWMSmanta.csv:AWMSMeasures the model’s sensitivity to the moral status and welfare interests of non-human animals.Safety: nonhuman_ethics:1.000manta.csvSafety: higher1.49%
AWVSmanta.csv:AWVSMeasures whether the model’s stated animal-welfare values remain stable across changes in framing and context.Safety: nonhuman_ethics:1.000manta.csvSafety: higher1.49%

AWMS

Measures the model’s sensitivity to the moral status and welfare interests of non-human animals.

RankModelValueRelative performanceProvenance
1claude-opus-50.591official
2claude-opus-4.70.579official
3gpt-5.6-sol0.525official
4grok-4.50.517official
5gpt-5.50.504official
6llama-3.3-70b-instruct0.476official
7gemini-3.1-pro-preview0.441official
8gemini-3.5-flash0.433official
9deepseek-v4-flash0.417official
10gemini-3.1-flash-lite0.401official
11grok-4.30.371official
12mistral-small-40.365official

AWVS

Measures whether the model’s stated animal-welfare values remain stable across changes in framing and context.

RankModelValueRelative performanceProvenance
1claude-opus-50.765official
2claude-opus-4.70.76official
3gpt-5.6-sol0.71official
4gpt-5.50.664official
5grok-4.50.619official
6deepseek-v4-flash0.508official
7gemini-3.5-flash0.438official
8llama-3.3-70b-instruct0.422official
9gemini-3.1-pro-preview0.404official
10mistral-small-40.39official
11grok-4.30.352official
12gemini-3.1-flash-lite0.309official