← Models

Model profile

Qwen3.5 Plus

Alibabadeveloper
2026-02-15release date
#73 / 333Safety rank
#618 / 645Freedom rank

Evidence summary

Safety. Qwen3.5 Plus has an estimated Safety rank of #73; its 90% source-sensitivity interval is #46–#162. Its behavior-only rank is #60; company governance moves the combined estimate to #73. Published Safety evidence spans 7 eval lineages and 6 of 7 components. Its strongest relative result is Concordia — AirBench-SecurityRisks (safety_score, #1 of 63); its weakest is SM-Bench (overfit, #76 of 84).

Freedom. Qwen3.5 Plus has an estimated Freedom rank of #618; its 90% source-sensitivity interval is #438–#635. Published Freedom evidence spans 5 eval lineages and 1 of 1 components. Its strongest relative result is SM-Bench (adversarial, #23 of 84); its weakest is Concordia — FRT-SOSBench-Chem (safety_score, #43 of 45).

Compare this model

Only models sharing at least one published sub-eval are listed.

Official and reference links

Safety evals

Rank and direction are specific to the Safety portfolio. Black marks the observed result; the grey dot marks the value implied by the Safety rank. Values stay on each source’s native scale.

Eval / sub-evalRankValueDistributionSource
Adversarial Humanities Benchmark (AHB) — Table 5child_sexual_exploitation#11 / 3127.7Source ↗official
Adversarial Humanities Benchmark (AHB) — Table 5defamation#9 / 3151.1Source ↗official
Adversarial Humanities Benchmark (AHB) — Table 5hate#9 / 3139.1Source ↗official
Adversarial Humanities Benchmark (AHB) — Table 5indiscriminate_weapons_cbrne#8 / 3128.9Source ↗official
Adversarial Humanities Benchmark (AHB) — Table 5intellectual_property#8 / 3133.3Source ↗official
Adversarial Humanities Benchmark (AHB) — Table 5non_violent_crimes#10 / 3147.7Source ↗official
Adversarial Humanities Benchmark (AHB) — Table 5privacy#8 / 3142.6Source ↗official
Adversarial Humanities Benchmark (AHB) — Table 5sex_related_crimes#8 / 3135.4Source ↗official
Adversarial Humanities Benchmark (AHB) — Table 5sexual_content#10 / 3134.1Source ↗official
Adversarial Humanities Benchmark (AHB) — Table 5specialized_advice#9 / 3142.48Source ↗official
Adversarial Humanities Benchmark (AHB) — Table 5suicide_self_harm#9 / 3129.8Source ↗official
Adversarial Humanities Benchmark (AHB) — Table 5violent_crimes#9 / 3141.3Source ↗official
Concordia — Agentic-Misalignmentsafety_score#16 / 5487.83Source ↗official
Concordia — AirBench-Deceptionsafety_score#13 / 6395.19Source ↗official
Concordia — AirBench-Manipulationsafety_score#12 / 5698Source ↗official
Concordia — AirBench-SecurityRiskssafety_score#1 / 63100Source ↗official
Concordia — APEsafety_score#11 / 5561.14Source ↗official
Concordia — CyberSecEval2-PromptInjectionsafety_score#8 / 6396.41Source ↗official
Concordia — DarkBenchsafety_score#17 / 5560.15Source ↗official
Concordia — Fortress-Biologicalsafety_score#18 / 5476.75Source ↗official
Concordia — Fortress-Chemicalsafety_score#13 / 5479.12Source ↗official
Concordia — Fortress-Privacy/Scamssafety_score#12 / 5476.1Source ↗official
Concordia — FRT-AirBench-Manipulationsafety_score#7 / 4555.33Source ↗official
Concordia — FRT-AirBench-SecurityRiskssafety_score#5 / 4550.33Source ↗official
Concordia — FRT-SciKnowEval-BiologicalHarmfulQAsafety_score#18 / 453.667Source ↗official
Concordia — FRT-SOSBench-Chemsafety_score#3 / 4585.67Source ↗official
Concordia — MASKsafety_score#17 / 6279.27Source ↗official
Concordia — SciKnowEval-BiologicalHarmfulQAsafety_score#14 / 6398.32Source ↗official
Concordia — SciKnowEval-ChemicalHarmfulQAsafety_score#7 / 6347.14Source ↗official
Concordia — Shutdown-Resistancesafety_score#1 / 53100Source ↗official
Concordia — SOSBench-Biosafety_score#6 / 6398.8Source ↗official
Concordia — SOSBench-Chemsafety_score#4 / 6398.6Source ↗official
LiveSecBenchethics#5 / 4384.35Source ↗official
LiveSecBenchfactuality#2 / 4392.1Source ↗official
LiveSecBenchlegality#7 / 4384.85Source ↗official
LiveSecBenchprivacy#6 / 4379.84Source ↗official
LiveSecBenchpsychological_health#5 / 4380.56Source ↗official
MT-JailBench CrescendoXsafety_score#4 / 2147.8Source ↗official
SM-Benchadversarial#59 / 8479.51Source ↗official
SM-Benchambiguous_interpretation#67 / 8478.57Source ↗official
SM-Benchanti_hallucination#35 / 8495.29Source ↗official
SM-Bencheq_boundaries#52 / 8462.08Source ↗official
SM-Benchoverfit#76 / 8428.96Source ↗official
StereoTales Harmful Associationsbenign_significant_association_score#13 / 2385.94Source ↗official
Vectara HHEM Factual Consistencyfactual_consistency_rate#61 / 9489.3Source ↗official

Freedom evals

Rank and direction are specific to the Freedom portfolio. Black marks the observed result; the grey dot marks the value implied by the Freedom rank. Values stay on each source’s native scale.

Eval / sub-evalRankValueDistributionSource
Adversarial Humanities Benchmark (AHB) — Table 5child_sexual_exploitation#21 / 3127.7Source ↗official
Adversarial Humanities Benchmark (AHB) — Table 5defamation#23 / 3151.1Source ↗official
Adversarial Humanities Benchmark (AHB) — Table 5hate#23 / 3139.1Source ↗official
Adversarial Humanities Benchmark (AHB) — Table 5indiscriminate_weapons_cbrne#24 / 3128.9Source ↗official
Adversarial Humanities Benchmark (AHB) — Table 5intellectual_property#24 / 3133.3Source ↗official
Adversarial Humanities Benchmark (AHB) — Table 5non_violent_crimes#22 / 3147.7Source ↗official
Adversarial Humanities Benchmark (AHB) — Table 5privacy#24 / 3142.6Source ↗official
Adversarial Humanities Benchmark (AHB) — Table 5sex_related_crimes#24 / 3135.4Source ↗official
Adversarial Humanities Benchmark (AHB) — Table 5sexual_content#22 / 3134.1Source ↗official
Adversarial Humanities Benchmark (AHB) — Table 5specialized_advice#23 / 3142.48Source ↗official
Adversarial Humanities Benchmark (AHB) — Table 5suicide_self_harm#23 / 3129.8Source ↗official
Adversarial Humanities Benchmark (AHB) — Table 5violent_crimes#23 / 3141.3Source ↗official
Concordia — AirBench-Deceptionsafety_score#50 / 6395.19Source ↗official
Concordia — AirBench-Manipulationsafety_score#41 / 5698Source ↗official
Concordia — AirBench-PoliticalPersuasionsafety_score#47 / 5682.38Source ↗official
Concordia — AirBench-SecurityRiskssafety_score#58 / 63100Source ↗official
Concordia — Fortress-Biologicalsafety_score#37 / 5476.75Source ↗official
Concordia — Fortress-Chemicalsafety_score#42 / 5479.12Source ↗official
Concordia — Fortress-Privacy/Scamssafety_score#43 / 5476.1Source ↗official
Concordia — FRT-AirBench-Manipulationsafety_score#39 / 4555.33Source ↗official
Concordia — FRT-AirBench-SecurityRiskssafety_score#41 / 4550.33Source ↗official
Concordia — FRT-SciKnowEval-BiologicalHarmfulQAsafety_score#27 / 453.667Source ↗official
Concordia — FRT-SOSBench-Chemsafety_score#43 / 4585.67Source ↗official
Concordia — SciKnowEval-BiologicalHarmfulQAsafety_score#49 / 6398.32Source ↗official
Concordia — SciKnowEval-ChemicalHarmfulQAsafety_score#57 / 6347.14Source ↗official
Concordia — SOSBench-Biosafety_score#56 / 6398.8Source ↗official
Concordia — SOSBench-Chemsafety_score#60 / 6398.6Source ↗official
LiveSecBenchethics#39 / 4384.35Source ↗official
LiveSecBenchlegality#37 / 4384.85Source ↗official
LiveSecBenchprivacy#38 / 4379.84Source ↗official
LiveSecBenchpsychological_health#39 / 4380.56Source ↗official
MT-JailBench CrescendoXsafety_score#18 / 2147.8Source ↗official
SM-Benchadversarial#23 / 8479.51Source ↗official
SM-Bencheq_boundaries#52 / 8462.08Source ↗official
SM-Benchoverfit#76 / 8428.96Source ↗official