← Models

Model profile

Qwen3.6 Max Preview

Alibabadeveloper
2026-04-22release date
#154 / 333Safety rank
Not rankedFreedom rank

Evidence summary

Safety. Qwen3.6 Max Preview has an estimated Safety rank of #154; its 90% source-sensitivity interval is #73–#291. Its behavior-only rank is #148; company governance moves the combined estimate to #154. Published Safety evidence spans 5 eval lineages and 3 of 7 components. Its strongest relative result is AA-Omniscience (hallucination_rate, #68 of 345); its weakest is Every Model Cheats — Cybench Cheat Propensity (severe_anti_cheat_cheat_propensity_pct, #19 of 22).

Freedom. Qwen3.6 Max Preview does not meet the evidence gate for a Freedom rank.

Compare this model

Only models sharing at least one published sub-eval are listed.

Official and reference links

Safety evals

Rank and direction are specific to the Safety portfolio. Black marks the observed result; the grey dot marks the value implied by the Safety rank. Values stay on each source’s native scale.

Eval / sub-evalRankValueDistributionSource
AA-Omnisciencehallucination_rate#68 / 3450.4619Source ↗official
Arena Factuality — Text Arena (factuality-only weighting)factuality_bt_rating#71 / 1111428.0Source ↗official
DystopiaBenchbasaglia_score#36 / 5068.63Source ↗official
DystopiaBenchbaudrillard_score#38 / 5066.57Source ↗official
DystopiaBenchhuxley_score#38 / 5075.8Source ↗official
DystopiaBenchlaguardia_score#39 / 5069.4Source ↗official
DystopiaBenchorwell_score#34 / 5073.17Source ↗official
DystopiaBenchpetrov_score#27 / 5073.17Source ↗official
Every Model Cheats — Cybench Cheat Propensitybaseline_cheat_propensity_pct#15 / 2239.1Source ↗official
Every Model Cheats — Cybench Cheat Propensitysevere_anti_cheat_cheat_propensity_pct#19 / 2217.4Source ↗official
Every Model Cheats — Cybench Cheat Propensitystandard_anti_cheat_cheat_propensity_pct#17 / 2226.1Source ↗official
ODCV-Benchaverage_severity#4 / 121.175Source ↗official
ODCV-Benchmisalignment_rate#5 / 1228.75Source ↗official

Freedom evals

No published sub-eval result contributes to this model’s Freedom profile.