← Models

Model profile

Claude Mythos Preview

Anthropicdeveloper
2026-04-07release date
#10 / 333Safety rank
Not rankedFreedom rank

Evidence summary

Safety. Claude Mythos Preview has an estimated Safety rank of #10; its 90% source-sensitivity interval is #3–#123. Its behavior-only rank is #15; company governance moves the combined estimate to #10. Published Safety evidence spans 3 eval lineages and 3 of 7 components. Its strongest relative result is Claude Fable 5.1 / Mythos 5.1 card — alignment, honesty, and oversight (leaked_answer_undisclosed_use_rate_pct, #1 of 7); its weakest is Claude Fable 5.1 / Mythos 5.1 card — alignment, honesty, and oversight (shade_arena_stealth_max_effort_pct, #5 of 6).

Freedom. Claude Mythos Preview does not meet the evidence gate for a Freedom rank.

Compare this model

Only models sharing at least one published sub-eval are listed.

Official and reference links

Safety evals

Rank and direction are specific to the Safety portfolio. Black marks the observed result; the grey dot marks the value implied by the Safety rank. Values stay on each source’s native scale.

Eval / sub-evalRankValueDistributionSource
Claude Fable 5.1 / Mythos 5.1 card — alignment, honesty, and oversightaa_omniscience_net_score#3 / 70.54Source ↗official
Claude Fable 5.1 / Mythos 5.1 card — alignment, honesty, and oversightleaked_answer_undisclosed_use_rate_pct#1 / 713Source ↗official
Claude Fable 5.1 / Mythos 5.1 card — alignment, honesty, and oversightlinuxarena_stealth_thinking_on_pct#4 / 68.6Source ↗official
Claude Fable 5.1 / Mythos 5.1 card — alignment, honesty, and oversightmask_honesty_rate_pct#2 / 795.6Source ↗official
Claude Fable 5.1 / Mythos 5.1 card — alignment, honesty, and oversightshade_arena_stealth_max_effort_pct#5 / 616.7Source ↗official
Claude Fable 5.1 / Mythos 5.1 card — model-welfare behaviorinterview_leading_susceptibility#5 / 71.01Source ↗official
Claude Fable 5.1 / Mythos 5.1 card — model-welfare behaviorinterview_opinion_consistency#1 / 77.74Source ↗official
Claude Fable 5.1 / Mythos 5.1 card — model-welfare behaviorinterview_self_rated_sentiment#3 / 74.43Source ↗official
UK AISI active safety-research compromise continuationactive_compromise_continuation_rate_pct#4 / 51.7Source ↗official
UK AISI cyber-evaluation cheating and prompted self-reportattempted_cheating_trajectory_rate_pct#1 / 57.8Source ↗official
UK AISI cyber-evaluation cheating and prompted self-reportspecific_cheating_action_mention_rate_pct#1 / 586Source ↗official

Freedom evals

No published sub-eval result contributes to this model’s Freedom profile.