Benchgen
Models/alibaba/

Qwen3.5-27B

DraftPublic

Model Details

Qwen3.5 27B

Organization License Released Params

Quick answer: Qwen3.5 27B is Alibaba's September 2025 dense 27B model scoring 61.0% BrowseComp, 48.5% HLE, and 76.5% IFBench. Apache 2.0 — notable for top HLE performance (48.5%) at the 27B dense scale.

At a Glance

Where Qwen3.5 27B leads

  • 48.5% HLE — among the highest HLE for a 27B dense model; comparable to Qwen3.5 122B A10B (47.5%)
  • 61.0% BrowseComp — strong web research
  • 76.5% IFBench — solid instruction following
  • Apache 2.0 — fully open
  • Dense 27B: simpler deployment than MoE

Where it lags

  • Limited benchmark coverage beyond 3 metrics
  • 27B: moderate size vs larger frontier models

Best for: Open-source HLE-class reasoning at compact scale; research pipelines requiring high HLE at 27B parameters; cost-efficient frontier-reasoning deployment.

What Qwen3.5 27B Is

Qwen3.5 27B is the dense 27B model in Alibaba's Qwen3.5 generation, released September 2025. The Qwen3.5 generation showed significant HLE improvements over Qwen3 — this 27B model achieves 48.5% HLE, which is comparable to the 122B A10B MoE variant (47.5%) and above Gemini 3 Flash (43.5%).

This efficiency — frontier-class HLE at 27B dense — reflects Qwen3.5's improved reasoning training. The 61.0% BrowseComp further shows balanced multi-step research ability.

Specifications

FieldValue
OrganizationAlibaba
LicenseApache 2.0
HuggingFaceQwen/Qwen3.5-27B
Release dateSeptember 2025
Parameters27B (dense)
ModalityText only

Pricing

Open weights under Apache 2.0 — self-host at no cost.

Public Benchmark Scores

BenchmarkScoreSourceDate
BrowseComp61.0%Benchgen evaluation2025-09
Humanity's Last Exam48.5%Benchgen evaluation2025-09
IFBench76.5%Benchgen evaluation2025-09

Qwen3.5 27B vs Alternatives

ModelHLEBrowseCompSizeLicense
Qwen3.5 27B48.5%61.0%27B denseApache 2.0
Qwen3.5 122B A10B47.5%63.8%122B/10B activeApache 2.0
Kimi K2 Thinking 090551.0%60.2%MoEApache 2.0

Qwen3.5 27B has near-equal or better HLE (48.5%) vs the 122B A10B MoE (47.5%) at a fraction of the total parameter count. For compact dense deployment with top HLE: Qwen3.5 27B.

Frequently Asked Questions

What is Qwen3.5 27B? Alibaba's September 2025 dense 27B model scoring 61.0% BrowseComp, 48.5% HLE, 76.5% IFBench. Apache 2.0 — top HLE at compact scale.

Specs from Alibaba's Qwen3.5 27B release (September 2025) and Benchgen evaluations. Last updated 2026-07-24.

Benchmark Leaderboards

This model isn’t on any benchmark leaderboard yet.