Quick answer: Qwen3.5 27B is Alibaba's September 2025 dense 27B model scoring 61.0% BrowseComp, 48.5% HLE, and 76.5% IFBench. Apache 2.0 — notable for top HLE performance (48.5%) at the 27B dense scale.
Where Qwen3.5 27B leads
Where it lags
Best for: Open-source HLE-class reasoning at compact scale; research pipelines requiring high HLE at 27B parameters; cost-efficient frontier-reasoning deployment.
Qwen3.5 27B is the dense 27B model in Alibaba's Qwen3.5 generation, released September 2025. The Qwen3.5 generation showed significant HLE improvements over Qwen3 — this 27B model achieves 48.5% HLE, which is comparable to the 122B A10B MoE variant (47.5%) and above Gemini 3 Flash (43.5%).
This efficiency — frontier-class HLE at 27B dense — reflects Qwen3.5's improved reasoning training. The 61.0% BrowseComp further shows balanced multi-step research ability.
| Field | Value |
|---|---|
| Organization | Alibaba |
| License | Apache 2.0 |
| HuggingFace | Qwen/Qwen3.5-27B |
| Release date | September 2025 |
| Parameters | 27B (dense) |
| Modality | Text only |
Open weights under Apache 2.0 — self-host at no cost.
| Benchmark | Score | Source | Date |
|---|---|---|---|
| BrowseComp | 61.0% | Benchgen evaluation | 2025-09 |
| Humanity's Last Exam | 48.5% | Benchgen evaluation | 2025-09 |
| IFBench | 76.5% | Benchgen evaluation | 2025-09 |
| Model | HLE | BrowseComp | Size | License |
|---|---|---|---|---|
| Qwen3.5 27B | 48.5% | 61.0% | 27B dense | Apache 2.0 |
| Qwen3.5 122B A10B | 47.5% | 63.8% | 122B/10B active | Apache 2.0 |
| Kimi K2 Thinking 0905 | 51.0% | 60.2% | MoE | Apache 2.0 |
Qwen3.5 27B has near-equal or better HLE (48.5%) vs the 122B A10B MoE (47.5%) at a fraction of the total parameter count. For compact dense deployment with top HLE: Qwen3.5 27B.
Specs from Alibaba's Qwen3.5 27B release (September 2025) and Benchgen evaluations. Last updated 2026-07-24.
This model isn’t on any benchmark leaderboard yet.