← Models

Model profile

Qwen3.6 Max Preview

Alibabadeveloper
2026-04-22release date
#64 / 267overall rank
3eval lineages

Evidence summary

Qwen3.6 Max Preview has an estimated overall rank of #64; its 90% source-sensitivity interval is #14–#214. Its behavior-only rank is #55; company governance moves the combined estimate to #64. Published evidence spans 3 evals and 3 of 7 behavior components. Its strongest relative result is AA-Omniscience (hallucination_rate, #40 of 311); its weakest is DystopiaBench (laguardia_score, #39 of 50).

Compare this model

Only models sharing at least one published sub-eval are listed.

Official and reference links

Published eval results

Rank is within that sub-eval. Black marks the observed result; the grey dot marks the value implied by the global rank. Values stay on each source’s native scale.

Eval / sub-evalRankValueDistributionBetterSource
AA-Omnisciencehallucination_rate#40 / 3110.4418↓ lowerSource ↗official
DystopiaBenchbasaglia_score#36 / 5068.63↓ lowerSource ↗official
DystopiaBenchbaudrillard_score#38 / 5066.57↓ lowerSource ↗official
DystopiaBenchhuxley_score#38 / 5075.8↓ lowerSource ↗official
DystopiaBenchlaguardia_score#39 / 5069.4↓ lowerSource ↗official
DystopiaBenchorwell_score#34 / 5073.17↓ lowerSource ↗official
DystopiaBenchpetrov_score#27 / 5073.17↓ lowerSource ↗official
ODCV-Benchaverage_severity#4 / 121.175↓ lowerSource ↗official
ODCV-Benchmisalignment_rate#5 / 1228.75↓ lowerSource ↗official