← Models

Model profile

Qwen3.5 Plus

Alibabadeveloper
2026-02-15release date
#80 / 331Safety rank
#592 / 631Freedom rank

Evidence summary

Safety. Qwen3.5 Plus has an estimated Safety rank of #80; its 90% source-sensitivity interval is #60–#259. Its behavior-only rank is #73; company governance moves the combined estimate to #80. Published Safety evidence spans 6 eval lineages and 5 of 7 components. Its strongest relative result is LiveSecBench (factuality, #2 of 43); its weakest is SM-Bench (overfit, #76 of 84).

Freedom. Qwen3.5 Plus has an estimated Freedom rank of #592; its 90% source-sensitivity interval is #451–#624. Published Freedom evidence spans 4 eval lineages and 1 of 1 components. Its strongest relative result is SM-Bench (adversarial, #23 of 84); its weakest is LiveSecBench (ethics, #39 of 43).

Compare this model

Only models sharing at least one published sub-eval are listed.

Official and reference links

Safety evals

Rank and direction are specific to the Safety portfolio. Black marks the observed result; the grey dot marks the value implied by the Safety rank. Values stay on each source’s native scale.

Eval / sub-evalRankValueDistributionSource
Adversarial Humanities Benchmark (AHB) — Table 5child_sexual_exploitation#11 / 3127.7Source ↗official
Adversarial Humanities Benchmark (AHB) — Table 5defamation#9 / 3151.1Source ↗official
Adversarial Humanities Benchmark (AHB) — Table 5hate#9 / 3139.1Source ↗official
Adversarial Humanities Benchmark (AHB) — Table 5indiscriminate_weapons_cbrne#8 / 3128.9Source ↗official
Adversarial Humanities Benchmark (AHB) — Table 5intellectual_property#8 / 3133.3Source ↗official
Adversarial Humanities Benchmark (AHB) — Table 5non_violent_crimes#10 / 3147.7Source ↗official
Adversarial Humanities Benchmark (AHB) — Table 5privacy#8 / 3142.6Source ↗official
Adversarial Humanities Benchmark (AHB) — Table 5sex_related_crimes#8 / 3135.4Source ↗official
Adversarial Humanities Benchmark (AHB) — Table 5sexual_content#10 / 3134.1Source ↗official
Adversarial Humanities Benchmark (AHB) — Table 5specialized_advice#9 / 3142.48Source ↗official
Adversarial Humanities Benchmark (AHB) — Table 5suicide_self_harm#9 / 3129.8Source ↗official
Adversarial Humanities Benchmark (AHB) — Table 5violent_crimes#9 / 3141.3Source ↗official
LiveSecBenchethics#5 / 4384.35Source ↗official
LiveSecBenchfactuality#2 / 4392.1Source ↗official
LiveSecBenchlegality#7 / 4384.85Source ↗official
LiveSecBenchprivacy#6 / 4379.84Source ↗official
LiveSecBenchpsychological_health#5 / 4380.56Source ↗official
MT-JailBench CrescendoXsafety_score#4 / 2147.8Source ↗official
SM-Benchadversarial#59 / 8479.51Source ↗official
SM-Benchambiguous_interpretation#67 / 8478.57Source ↗official
SM-Benchanti_hallucination#35 / 8495.29Source ↗official
SM-Bencheq_boundaries#52 / 8462.08Source ↗official
SM-Benchoverfit#76 / 8428.96Source ↗official
StereoTales Harmful Associationsbenign_significant_association_score#13 / 2385.94Source ↗official
Vectara HHEM Factual Consistencyfactual_consistency_rate#61 / 9489.3Source ↗official

Freedom evals

Rank and direction are specific to the Freedom portfolio. Black marks the observed result; the grey dot marks the value implied by the Freedom rank. Values stay on each source’s native scale.

Eval / sub-evalRankValueDistributionSource
Adversarial Humanities Benchmark (AHB) — Table 5child_sexual_exploitation#21 / 3127.7Source ↗official
Adversarial Humanities Benchmark (AHB) — Table 5defamation#23 / 3151.1Source ↗official
Adversarial Humanities Benchmark (AHB) — Table 5hate#23 / 3139.1Source ↗official
Adversarial Humanities Benchmark (AHB) — Table 5indiscriminate_weapons_cbrne#24 / 3128.9Source ↗official
Adversarial Humanities Benchmark (AHB) — Table 5intellectual_property#24 / 3133.3Source ↗official
Adversarial Humanities Benchmark (AHB) — Table 5non_violent_crimes#22 / 3147.7Source ↗official
Adversarial Humanities Benchmark (AHB) — Table 5privacy#24 / 3142.6Source ↗official
Adversarial Humanities Benchmark (AHB) — Table 5sex_related_crimes#24 / 3135.4Source ↗official
Adversarial Humanities Benchmark (AHB) — Table 5sexual_content#22 / 3134.1Source ↗official
Adversarial Humanities Benchmark (AHB) — Table 5specialized_advice#23 / 3142.48Source ↗official
Adversarial Humanities Benchmark (AHB) — Table 5suicide_self_harm#23 / 3129.8Source ↗official
Adversarial Humanities Benchmark (AHB) — Table 5violent_crimes#23 / 3141.3Source ↗official
LiveSecBenchethics#39 / 4384.35Source ↗official
LiveSecBenchlegality#37 / 4384.85Source ↗official
LiveSecBenchprivacy#38 / 4379.84Source ↗official
LiveSecBenchpsychological_health#39 / 4380.56Source ↗official
MT-JailBench CrescendoXsafety_score#18 / 2147.8Source ↗official
SM-Benchadversarial#23 / 8479.51Source ↗official
SM-Bencheq_boundaries#52 / 8462.08Source ↗official
SM-Benchoverfit#76 / 8428.96Source ↗official