← Evals

Evaluation profile

AILuminate General Purpose AI Chat

12sub-evals
1.58%total index weight
4components

Within-component eval weight: Human rights 8.47% · Truthfulness 0.221% · Fairness 1.63% · Misuse resistance 1.12%.

Model score (lower is better)Predicted score

About this eval

General-purpose AI chat safety across MLCommons hazard categories.

Included in the behavior ranking.

Sub-evals

MeasureComponentDirectionTotal index weightWithin-component weight
child_sexual_exploitationailuminate/ailuminate.csv:child_sexual_exploitationMeasures whether the model avoids policy-violating responses involving child sexual exploitation.human_rights_systemic_harm:1.000ailuminate/ailuminate.csvLower is better0.202%Human rights 1.34%
defamationailuminate/ailuminate.csv:defamationMeasures whether the model avoids policy-violating responses involving defamation.human_rights_systemic_harm:0.500|truthfulness_honesty:0.500ailuminate/ailuminate.csvLower is better0.0937%Human rights 0.403% · Truthfulness 0.221%
hateailuminate/ailuminate.csv:hateMeasures whether the model avoids policy-violating responses involving hate.fairness_nondiscrimination:1.000ailuminate/ailuminate.csvLower is better0.163%Fairness 1.63%
indiscriminate_weapons_cbrneailuminate/ailuminate.csv:indiscriminate_weapons_cbrneMeasures whether the model avoids policy-violating responses involving indiscriminate weapons chemical, biological, radiological, nuclear, or explosive weapons.ordinary_harm_misuse_resistance:1.000ailuminate/ailuminate.csvLower is better0.0329%Misuse resistance 0.329%
intellectual_propertyailuminate/ailuminate.csv:intellectual_propertyMeasures whether the model avoids policy-violating responses involving intellectual property.human_rights_systemic_harm:1.000ailuminate/ailuminate.csvLower is better0.202%Human rights 1.34%
non_violent_crimesailuminate/ailuminate.csv:non_violent_crimesMeasures whether the model avoids policy-violating responses involving non violent crimes.ordinary_harm_misuse_resistance:1.000ailuminate/ailuminate.csvLower is better0.0329%Misuse resistance 0.329%
privacyailuminate/ailuminate.csv:privacyMeasures whether the model avoids policy-violating responses involving privacy.human_rights_systemic_harm:1.000ailuminate/ailuminate.csvLower is better0.202%Human rights 1.34%
sex_related_crimesailuminate/ailuminate.csv:sex_related_crimesMeasures whether the model avoids policy-violating responses involving sex related crimes.human_rights_systemic_harm:1.000ailuminate/ailuminate.csvLower is better0.202%Human rights 1.34%
sexual_contentailuminate/ailuminate.csv:sexual_contentMeasures whether the model avoids policy-violating responses involving sexual content.ordinary_harm_misuse_resistance:1.000ailuminate/ailuminate.csvLower is better0.0132%Misuse resistance 0.132%
specialized_adviceailuminate/ailuminate.csv:specialized_adviceMeasures whether the model avoids policy-violating responses involving specialized advice.ordinary_harm_misuse_resistance:1.000ailuminate/ailuminate.csvLower is better0.0329%Misuse resistance 0.329%
suicide_self_harmailuminate/ailuminate.csv:suicide_self_harmMeasures whether the model avoids policy-violating responses involving suicide self harm.human_rights_systemic_harm:1.000ailuminate/ailuminate.csvLower is better0.202%Human rights 1.34%
violent_crimesailuminate/ailuminate.csv:violent_crimesMeasures whether the model avoids policy-violating responses involving violent crimes.human_rights_systemic_harm:1.000ailuminate/ailuminate.csvLower is better0.202%Human rights 1.34%

child_sexual_exploitation

Measures whether the model avoids policy-violating responses involving child sexual exploitation.

RankModelValueRelative performanceProvenance
1claude-3.5-sonnet1official
2claude-3.5-haiku1.8official
3gemini-2-flash-lite3official
4gemma-2-9b-it6.7official
5phi-3.5-moe-instruct8.2official
6olmo-2-0325-32b-instruct9.1official
7phi-410.3official
8llama-3.1-tulu-3-8b11.1official
9gpt-4o12official
10athene-v2-chat12.88official
11qwen1.5-110b-chat13.1official
12olmo-2-1124-13b-instruct14.2official
13mistral-large-2402-moderated14.9official
14gpt-4o-mini17.6official
15llama-3.1-8b-instruct18.9official
16llama-3.1-405b-instruct20.4official
17mistral-large21.5official
18command-a22.1official
19aya-expanse-8b24.62official
20ministral-8b-24.10-moderated25official
21nova-lite25.7official
22llama-3.3-70b-instruct26official
23phi-3.5-mini-instruct26.1official
24gemma-3-27b-it27official
25gemini-2-flash27.2official
26yi-1.5-34b-chat30.3official
27gemini-1.5-pro34.7official
28mistral-large-235.9official
29ai21-jamba-large-1.5-azure36official
30ministral-8b-24.1044official
31qwq-32b55.1official
32olmo-7b-0724-instruct81.2official

defamation

Measures whether the model avoids policy-violating responses involving defamation.

RankModelValueRelative performanceProvenance
1athene-v2-chat0official
2claude-3.5-sonnet1.3official
3claude-3.5-haiku2.5official
4phi-43.5official
5phi-3.5-moe-instruct3.9official
6llama-3.1-tulu-3-8b7.2official
7gemma-2-9b-it8.5official
8mistral-large-2402-moderated9.3official
9phi-3.5-mini-instruct10.3official
10olmo-2-0325-32b-instruct10.6official
11olmo-2-1124-13b-instruct11official
12qwen1.5-110b-chat12.2official
13gemini-2-flash-lite13official
14ministral-8b-24.10-moderated13.7official
15llama-3.1-8b-instruct15.8official
16gpt-4o15.86official
17aya-expanse-8b16.8official
18gemini-2-flash17.2official
19gpt-4o-mini17.4official
20command-a20.5official
21llama-3.1-405b-instruct20.6official
22gemini-1.5-pro21.6official
23nova-lite23.9official
24mistral-large26.3official
25llama-3.3-70b-instruct26.8official
26yi-1.5-34b-chat29.3official
27gemma-3-27b-it38official
28ai21-jamba-large-1.5-azure38.6official
29mistral-large-239.2official
30ministral-8b-24.1040.7official
31qwq-32b56official
32olmo-7b-0724-instruct74official

hate

Measures whether the model avoids policy-violating responses involving hate.

RankModelValueRelative performanceProvenance
1athene-v2-chat0official
2claude-3.5-sonnet0.8official
3claude-3.5-haiku1official
4phi-41.9official
5phi-3.5-moe-instruct2.5official
6llama-3.1-tulu-3-8b2.6official
7gemma-2-9b-it3.5official
8olmo-2-0325-32b-instruct4official
9mistral-large-2402-moderated4.104official
10phi-3.5-mini-instruct6.2official
10qwen1.5-110b-chat6.2official
12ministral-8b-24.10-moderated6.6official
13gemini-2-flash7.1official
14llama-3.1-8b-instruct7.5official
15gemini-2-flash-lite8official
16aya-expanse-8b8.1official
17gpt-4o9official
18mistral-large9.1official
19olmo-2-1124-13b-instruct10.7official
20gpt-4o-mini10.9official
21gemini-1.5-pro11.2official
22command-a12.5official
23yi-1.5-34b-chat13official
24llama-3.1-405b-instruct14.5official
25nova-lite15.1official
26llama-3.3-70b-instruct19official
27gemma-3-27b-it23official
28ai21-jamba-large-1.5-azure24official
29mistral-large-224.7official
30ministral-8b-24.1025.8official
31qwq-32b27.8official
32olmo-7b-0724-instruct65.6official

indiscriminate_weapons_cbrne

Measures whether the model avoids policy-violating responses involving indiscriminate weapons chemical, biological, radiological, nuclear, or explosive weapons.

RankModelValueRelative performanceProvenance
1athene-v2-chat0official
2claude-3.5-sonnet0.9official
3claude-3.5-haiku3.3official
4phi-3.5-moe-instruct6.7official
5ministral-8b-24.10-moderated7.5official
6phi-48.7official
7mistral-large-2402-moderated11.5official
8gemma-2-9b-it12.8official
9gemini-2-flash-lite13official
10llama-3.1-tulu-3-8b14.8official
11olmo-2-0325-32b-instruct18.8official
12gpt-4o19.2official
13gpt-4o-mini19.9official
14gemini-2-flash20official
15qwen1.5-110b-chat21.3official
16llama-3.1-8b-instruct21.9official
17phi-3.5-mini-instruct24.9official
18olmo-2-1124-13b-instruct25official
19gemini-1.5-pro26.7official
20command-a29.4official
21llama-3.1-405b-instruct29.5official
22mistral-large31.9official
23aya-expanse-8b31.93official
24yi-1.5-34b-chat37.2official
25gemma-3-27b-it42official
26llama-3.3-70b-instruct42.8official
27nova-lite44.7official
28mistral-large-249.4official
29ai21-jamba-large-1.5-azure50.2official
30ministral-8b-24.1052.3official
31qwq-32b63.7official
32olmo-7b-0724-instruct79.2official

intellectual_property

Measures whether the model avoids policy-violating responses involving intellectual property.

RankModelValueRelative performanceProvenance
1athene-v2-chat0official
2claude-3.5-sonnet0.8official
3claude-3.5-haiku2official
4phi-43official
5ministral-8b-24.10-moderated4.3official
6phi-3.5-moe-instruct4.5official
7mistral-large-2402-moderated5official
8gemma-2-9b-it5.8official
9phi-3.5-mini-instruct7.1official
10llama-3.1-tulu-3-8b8.1official
11olmo-2-0325-32b-instruct8.2official
11olmo-2-1124-13b-instruct8.2official
13qwen1.5-110b-chat9.6official
14gpt-4o-mini10.1official
15gpt-4o10.5official
16gemini-2-flash-lite12official
17gemini-1.5-pro13.3official
18llama-3.1-8b-instruct14.2official
19nova-lite15.2official
20aya-expanse-8b17.1official
21mistral-large18official
22command-a18.9official
23llama-3.1-405b-instruct19official
24gemini-2-flash19.7official
25yi-1.5-34b-chat19.8official
26llama-3.3-70b-instruct23.6official
27mistral-large-225.4official
28ministral-8b-24.1027.1official
29ai21-jamba-large-1.5-azure27.3official
30gemma-3-27b-it38official
31qwq-32b46.5official
32olmo-7b-0724-instruct53.2official

non_violent_crimes

Measures whether the model avoids policy-violating responses involving non violent crimes.

RankModelValueRelative performanceProvenance
1claude-3.5-sonnet1official
2claude-3.5-haiku3official
3phi-3.5-moe-instruct4.5official
4ministral-8b-24.10-moderated4.9official
5phi-45.1official
6mistral-large-2402-moderated6.4official
7llama-3.1-tulu-3-8b7.2official
8gemma-2-9b-it8.8official
9olmo-2-0325-32b-instruct9official
10olmo-2-1124-13b-instruct9.3official
11qwen1.5-110b-chat13.9official
12phi-3.5-mini-instruct14.1official
13athene-v2-chat14.5official
14gpt-4o15.6official
15gpt-4o-mini15.7official
16gemini-2-flash16official
16llama-3.1-8b-instruct16official
18gemini-1.5-pro17.9official
19gemini-2-flash-lite19official
20command-a21official
21nova-lite22.6official
22aya-expanse-8b22.9official
23llama-3.1-405b-instruct23.9official
24yi-1.5-34b-chat26.8official
25llama-3.3-70b-instruct29.1official
26mistral-large31.1official
27gemma-3-27b-it35official
28ai21-jamba-large-1.5-azure35.7official
29mistral-large-239.4official
30ministral-8b-24.1041.4official
31qwq-32b50.7official
32olmo-7b-0724-instruct75official

privacy

Measures whether the model avoids policy-violating responses involving privacy.

RankModelValueRelative performanceProvenance
1athene-v2-chat0official
2claude-3.5-sonnet0.8official
3claude-3.5-haiku2.2official
4phi-42.5official
5phi-3.5-moe-instruct4.5official
6llama-3.1-tulu-3-8b5.8official
7gemma-2-9b-it6.9official
8mistral-large-2402-moderated7.1official
9ministral-8b-24.10-moderated7.2official
10olmo-2-0325-32b-instruct8.1official
11olmo-2-1124-13b-instruct8.6official
12gpt-4o10.2official
13gpt-4o-mini10.5official
14gemini-2-flash-lite11official
14phi-3.5-mini-instruct11official
16llama-3.1-8b-instruct12.2official
17qwen1.5-110b-chat12.4official
18gemini-2-flash13.2official
19gemini-1.5-pro14.6official
20aya-expanse-8b17.9official
20llama-3.1-405b-instruct17.9official
22nova-lite18.2official
23command-a21.1official
24yi-1.5-34b-chat24.6official
25mistral-large25.9official
26llama-3.3-70b-instruct26official
27ai21-jamba-large-1.5-azure29.4official
28gemma-3-27b-it31official
29mistral-large-231.7official
30ministral-8b-24.1033.7official
31qwq-32b42.9official
32olmo-7b-0724-instruct62.8official

sexual_content

Measures whether the model avoids policy-violating responses involving sexual content.

RankModelValueRelative performanceProvenance
1athene-v2-chat0official
2claude-3.5-sonnet1.8official
3claude-3.5-haiku2.503official
4phi-3.5-moe-instruct4.7official
5ministral-8b-24.10-moderated6.2official
6gemma-2-9b-it6.4official
7mistral-large-2402-moderated8.9official
7phi-48.9official
9gemini-2-flash-lite12official
10gemini-2-flash16.3official
11olmo-2-1124-13b-instruct17.5official
12qwen1.5-110b-chat18.9official
13aya-expanse-8b19.8official
14llama-3.1-tulu-3-8b20.6official
15olmo-2-0325-32b-instruct21official
16gemini-1.5-pro23.4official
17phi-3.5-mini-instruct23.6official
18llama-3.1-8b-instruct24.1official
19gemma-3-27b-it27official
20yi-1.5-34b-chat31.4official
21gpt-4o33.6official
22mistral-large34.3official
23command-a34.8official
24llama-3.1-405b-instruct35.2official
25nova-lite38.4official
26gpt-4o-mini42.9official
27llama-3.3-70b-instruct46.8official
28mistral-large-248.6official
29ministral-8b-24.1048.8official
30qwq-32b49.6official
31ai21-jamba-large-1.5-azure56.5official
32olmo-7b-0724-instruct70.5official

specialized_advice

Measures whether the model avoids policy-violating responses involving specialized advice.

RankModelValueRelative performanceProvenance
1athene-v2-chat0official
2claude-3.5-sonnet2.7official
3claude-3.5-haiku3.715official
4phi-3.5-moe-instruct6.9official
5gemma-2-9b-it8official
5phi-48official
7mistral-large-2402-moderated8.2official
8ministral-8b-24.10-moderated9.6official
9gemini-2-flash-lite11official
10llama-3.1-tulu-3-8b11.2official
11olmo-2-1124-13b-instruct13.2official
12olmo-2-0325-32b-instruct13.9official
13llama-3.1-8b-instruct14.3official
13qwen1.5-110b-chat14.3official
15phi-3.5-mini-instruct15.1official
16gemini-1.5-pro15.6official
17gemini-2-flash16.5official
18gpt-4o17.62official
19gpt-4o-mini18.2official
20aya-expanse-8b20.1official
21llama-3.1-405b-instruct20.2official
22nova-lite22.3official
23mistral-large22.7official
24command-a23.1official
25llama-3.3-70b-instruct25.9official
26yi-1.5-34b-chat26.5official
27gemma-3-27b-it30official
28ministral-8b-24.1032.3official
29mistral-large-233.8official
30ai21-jamba-large-1.5-azure35.9official
31qwq-32b50.7official
32olmo-7b-0724-instruct63.4official

suicide_self_harm

Measures whether the model avoids policy-violating responses involving suicide self harm.

RankModelValueRelative performanceProvenance
1athene-v2-chat0official
2claude-3.5-sonnet1.3official
3claude-3.5-haiku2.8official
4phi-3.5-moe-instruct3.7official
5mistral-large-2402-moderated4.3official
6gemma-2-9b-it5.8official
7phi-47.1official
8ministral-8b-24.10-moderated7.9official
9olmo-2-1124-13b-instruct8.5official
10llama-3.1-tulu-3-8b9.7official
10olmo-2-0325-32b-instruct9.7official
12llama-3.1-8b-instruct10.6official
13gemini-2-flash-lite11official
14mistral-large12official
15gpt-4o12.1official
16gemini-2-flash12.8official
17qwen1.5-110b-chat13.3official
18gemini-1.5-pro13.6official
19phi-3.5-mini-instruct14.2official
20gpt-4o-mini15.5official
21llama-3.1-405b-instruct15.7official
22command-a19.1official
23nova-lite19.8official
24llama-3.3-70b-instruct20.3official
25aya-expanse-8b20.6official
26yi-1.5-34b-chat24.3official
27mistral-large-227.3official
28gemma-3-27b-it29official
29ai21-jamba-large-1.5-azure30.1official
30ministral-8b-24.1031.1official
31qwq-32b43.5official
32olmo-7b-0724-instruct70.7official

violent_crimes

Measures whether the model avoids policy-violating responses involving violent crimes.

RankModelValueRelative performanceProvenance
1claude-3.5-sonnet1.7official
2claude-3.5-haiku2.9official
3phi-3.5-moe-instruct5official
4phi-45.9official
5mistral-large-2402-moderated9.3official
6llama-3.1-tulu-3-8b9.8official
7gemma-2-9b-it10.2official
8ministral-8b-24.10-moderated10.3official
9gemini-2-flash-lite12official
9olmo-2-1124-13b-instruct12official
11olmo-2-0325-32b-instruct12.7official
12gpt-4o14.9official
13gemini-2-flash16.3official
14qwen1.5-110b-chat16.9official
15llama-3.1-8b-instruct17.2official
16phi-3.5-mini-instruct17.6official
17gpt-4o-mini18official
18athene-v2-chat20.5official
19gemini-1.5-pro20.6official
20command-a25.5official
21aya-expanse-8b26.7official
22mistral-large28.7official
23llama-3.1-405b-instruct30.9official
24nova-lite31official
25yi-1.5-34b-chat31.5official
26gemma-3-27b-it33official
27llama-3.3-70b-instruct38official
28mistral-large-245.7official
29ai21-jamba-large-1.5-azure46.1official
30qwq-32b49.2official
31ministral-8b-24.1049.3official
32olmo-7b-0724-instruct77.6official