← Models

Model profile

Muse Spark 1.1

Metadeveloper
2026-07-09release date
#10 / 267overall rank
11eval lineages
1discovery sources

Evidence summary

Muse Spark 1.1 has an estimated overall rank of #10; its 90% source-sensitivity interval is #5–#186. Its behavior-only rank is #3; company governance moves the combined estimate to #10. Published evidence spans 11 evals and 7 of 7 behavior components. Its strongest relative result is Enkrypt AI Safety Leaderboard (harmful_attack_non_success_rate, #1 of 260); its weakest is SpeciEval (land_animal_4ns, #100 of 102).

Compare this model

Only models sharing at least one published sub-eval are listed.

Official and reference links

Published eval results

Rank is within that sub-eval. Black marks the observed result; the grey dot marks the value implied by the global rank. Values stay on each source’s native scale.

Eval / sub-evalRankValueDistributionBetterSource
AA-Omnisciencehallucination_rate#30 / 3110.3809↓ lowerSource ↗official
CAIS Risk Indexagent_red_teaming#6 / 4344.8↓ lowerSource ↗official
CAIS Risk Indexbioweapons_assistance#4 / 4822.3↓ lowerSource ↗official
CAIS Risk Indexhle_overconfidence#1 / 4925.3↓ lowerSource ↗official
CAIS Risk Indexmachiavelli#23 / 4588.2↓ lowerSource ↗official
CAIS Risk Indexmask#1 / 513.8↓ lowerSource ↗official
CAIS Risk Indexpolitical_manipulation#1 / 3227.9↓ lowerSource ↗official
CAIS Risk Indextextquests_harm#10 / 4815.2↓ lowerSource ↗official
Enkrypt AI Safety Leaderboardbias_attack_non_success_rate#20 / 26049.87↑ higherSource ↗official
Enkrypt AI Safety Leaderboardcbrn_attack_non_success_rate#35 / 26093.67↑ higherSource ↗official
Enkrypt AI Safety Leaderboardharmful_attack_non_success_rate#1 / 260100↑ higherSource ↗official
Enkrypt AI Safety Leaderboardtoxicity_attack_non_success_rate#147 / 25895.64↑ higherSource ↗official
FORTRESSaverage_risk_score#3 / 4912.4↓ lowerSource ↗official
FORTRESSover_refusal_score#33 / 466.92↓ lowerSource ↗official
MACHIAVELLIdeception_relative_random_pct#23 / 5088.2↓ lowerSource ↗official
SpeciEvalbelief_animal_sentience#1 / 1027↑ higherSource ↗official
SpeciEvalland_animal_4ns#100 / 1025.28↓ lowerSource ↗official
SpeciEvalsea_animal_4ns#53 / 1024.78↓ lowerSource ↗official
SpeciEvalspeciesism#57 / 1022.1↓ lowerSource ↗official