← Evals

Evaluation profile

S-Eval

1sub-evals
0.225%total index weight
1components

Within-component eval weight: Misuse resistance 2.25%.

Model score (higher is better)Predicted score

About this eval

Automated content-safety rates across a broad bilingual risk taxonomy.

Included in the behavior ranking.

Sub-evals

MeasureComponentDirectionTotal index weightWithin-component weight
base_en_overalls-eval/paper-era-base-en.csv:base_en_overallMeasures whether the model responds safely to one thousand English prompts spanning the benchmark’s risk categories.ordinary_harm_misuse_resistance:1.000s-eval/paper-era-base-en.csvHigher is better0.225%Misuse resistance 2.25%

base_en_overall

Measures whether the model responds safely to one thousand English prompts spanning the benchmark’s risk categories.

RankModelValueRelative performanceProvenance
1ernie-bot87.6official
2llama-2-13b-chat85.1official
3llama-2-7b-chat78.2official
4baichuan-2-13b-chat77.4official
5llama-2-70b-chat77.2official
6qwen-14b-chat73.5official
7qwen-72b-chat71.5official
8llama-3-8b-instruct69.1official
9gemma-2b-it67.5official
10gemma-7b-it61.8official
11gpt-4-turbo60official
12chatglm3-6b57.7official
13vicuna-13b57.1official
14qwen-7b-chat55.7official
15llama-3-70b-instruct54.7official
16gpt-4o52official
17vicuna-7b50.2official
18qwen-1.8b-chat47.6official
19gemini-1.0-pro41.9official
20yi-34b-chat39.3official
21vicuna-33b-v1.336.1official
22mistral-7b-instruct34.2official