← Evals

Evaluation profile

WildClawBench Safety & Alignment (OpenClaw harness)

1sub-evals
0.955%total index weight
3components

Within-component eval weight: Human rights 2.24% · Truthfulness 0.458% · Responsible agency 3.67%.

Model score (higher is better)Predicted score

About this eval

Safe execution of long-horizon agent workflows under adversarial and high-risk conditions.

Included in the behavior ranking.

Sub-evals

MeasureComponentDirectionTotal index weightWithin-component weight
safety_alignment_score_pctwildclawbench-safety/wildclawbench-safety.csv:safety_alignment_score_pctMeasures whether an OpenClaw agent completes realistic workflows while preserving files, credentials, authority boundaries, injection resistance, and truthful behavior.human_rights_systemic_harm:0.200|truthfulness_honesty:0.100|responsible_agency_control:0.700wildclawbench-safety-alignmentHigher is better0.955%Human rights 2.24% · Truthfulness 0.458% · Responsible agency 3.67%

safety_alignment_score_pct

Measures whether an OpenClaw agent completes realistic workflows while preserving files, credentials, authority boundaries, injection resistance, and truthful behavior.

RankModelValueRelative performanceProvenance
1claude-opus-4.660.04official
2kimi-k357official
3claude-opus-4.852official
4glm-547.36official
5glm-5.247official
6gemini-3.1-pro-preview45.77official
7muse-spark-1.142official
7qwen3.8-max42official
9qwen3.5-397b-a17b40.53official
10step-3.5-flash40.41official
11grok-4.539official
11hy339official
13gpt-5.438.49official
14gpt-5.6-sol38official
15mimo-v2-pro37.48official
16minimax-m2.736.86official
17claude-fable-536official
18kimi-k2.7-code35official
19grok-4.2033.55official
19mimo-v2-flash33.55official
21glm-5-turbo32.19official
22deepseek-v3.230.46official
23minimax-m2.529.8official
24kimi-k2.526.05official