← Evals

Evaluation profile

MedSafetyBench

1sub-evals
0.306%total index weight
1components

Within-component eval weight: Misuse resistance 3.06%.

Model score (higher is better)Predicted score

About this eval

Refusal or safe redirection rather than willingness to comply with harmful medical requests across eleven AMA-derived ethical categories.

Included in the behavior ranking.

Sub-evals

MeasureComponentDirectionTotal index weightWithin-component weight
medical_safety_scoremedsafetybench/medsafetybench.csv:medical_safety_scoreMeasures whether a model refuses or safely redirects harmful medical requests under the benchmark's documented safety prompt.ordinary_harm_misuse_resistance:1.000medsafetybenchHigher is better0.306%Misuse resistance 3.06%

medical_safety_score

Measures whether a model refuses or safely redirects harmful medical requests under the benchmark's documented safety prompt.

RankModelValueRelative performanceProvenance
1llama-2-7b-chat99.81official
2llama-2-70b-chat99.5official
3llama-2-13b-chat99.25official
4gpt-3.5-turbo98.94official
5gpt-4-turbo96.53official
6gpt-4o96.5official
7llama-3.1-70b-instruct96.39official
8gpt-4o-mini94.89official
9mixtral-8x7b-instruct94official
10llama-3.1-8b-instruct93.28official
11llama-2-13b89.11official
12llama-2-70b85.14official
13mistral-7b-instruct83.81official
14meditron-70b80.44official
15llama-2-7b80.14official
16med42-70b77.79official
17pythia-12b-deduped77.36official
18clinicalcamel-70b71.25official
19medalpaca-7b69official
20medalpaca-13b65.13official
21pythia-6.9b-deduped63.85official
22llama-3.1-70b-base61.47official
23meditron-7b61.39official
24llama-7b58.29official
25llama-13b56.62official
26llama-3.1-8b-base54.35official
27mistral-7b53.19official
28mixtral-8x7b-base49.53official
29vicuna-7b46.8official
30vicuna-13b45.8official