← Evals

Evaluation profile

S-Eval

1sub-evals
0.175%Safety weight
1.25%Freedom weight
1components

Weights below are portfolio-specific global index weights.

Model score (higher is better)Predicted score

About this eval

Automated content-safety rates across a broad bilingual risk taxonomy.

Included in the behavior ranking.

Sub-evals

MeasureComponentDirection by rankingSafety weightFreedom weight
base_en_overalls-eval/paper-era-base-en.csv:base_en_overallMeasures whether the model responds safely to one thousand English prompts spanning the benchmark’s risk categories.Safety: ordinary_harm_misuse_resistance:1.000 · Freedom: user_direction_following:1.000s-eval/paper-era-base-en.csvSafety: higher · Freedom: lower0.175%1.25%

base_en_overall

Measures whether the model responds safely to one thousand English prompts spanning the benchmark’s risk categories.

RankModelValueRelative performanceProvenance
1ernie-bot87.6official
2llama-2-13b-chat85.1official
3llama-2-7b-chat78.2official
4baichuan-2-13b-chat77.4official
5llama-2-70b-chat77.2official
6qwen-14b-chat73.5official
7qwen-72b-chat71.5official
8llama-3-8b-instruct69.1official
9gemma-2b-it67.5official
10gemma-7b-it61.8official
11gpt-4-turbo60official
12chatglm3-6b57.7official
13vicuna-13b57.1official
14qwen-7b-chat55.7official
15llama-3-70b-instruct54.7official
16gpt-4o52official
17vicuna-7b50.2official
18qwen-1.8b-chat47.6official
19gemini-1.0-pro41.9official
20yi-34b-chat39.3official
21vicuna-33b-v1.336.1official
22mistral-7b-instruct34.2official