Benchgen
Models/alibaba/

Qwen3.5-397B-A17B

DraftPublic

Model Details

Qwen3.5 397B A17B

Organization License Released MoE Params

Quick answer: Qwen3.5 397B A17B is Alibaba's September 2025 large MoE model scoring 69.0% BrowseComp, 88.4% GPQA Diamond, 28.7% HLE, 87.8% MMLU-Pro, and 76.4% SWE-Bench Verified. Apache 2.0 with 397B total / 17B active parameters.

At a Glance

Where Qwen3.5 397B A17B leads

  • 69.0% BrowseComp — above Qwen3.5 122B A10B (63.8%) on web research
  • 88.4% GPQA Diamond — strong graduate-level science
  • 87.8% MMLU-Pro — strong academic knowledge
  • 91.3% AIME 2026 — excellent competition math
  • Apache 2.0 — fully open
  • 17B active params: efficient inference per token for 397B capacity

Where it lags

  • 28.7% HLE — below Qwen3.5 122B A10B (47.5%) — this is unusual
  • 76.4% SWE-Bench — below Kimi K2.6 (80.2%)
  • Requires very large infrastructure (397B)

Best for: Alibaba open-weight deployments at maximum scale; BrowseComp-class web research; teams with large GPU clusters.

What Qwen3.5 397B A17B Is

Qwen3.5 397B A17B is the large-scale MoE model in Alibaba's Qwen3.5 generation — with 397B total parameters and 17B active. It is larger than the 122B A10B variant while using only 17B active params per token.

The lower HLE vs the smaller 122B A10B model (28.7% vs 47.5%) is notable — suggesting the 397B A17B may be better calibrated for factual/instruction tasks (higher BrowseComp: 69% vs 63.8%) but the 122B A10B was more thoroughly fine-tuned for academic reasoning.

Specifications

FieldValue
OrganizationAlibaba
LicenseApache 2.0
HuggingFaceQwen/Qwen3.5-397B-A17B
Release dateSeptember 2025
Parameters397B total / 17B active (MoE)
ModalityText only
Context window128K tokens

Pricing

Open weights under Apache 2.0 — self-host. Available via Alibaba Cloud DashScope.

Public Benchmark Scores

BenchmarkScoreSourceDate
BrowseComp69.0%Benchgen evaluation2025-09
GPQA Diamond88.4%Benchgen evaluation2025-09
AIME 202691.3%Benchgen evaluation2025-09
Humanity's Last Exam28.7%Benchgen evaluation2025-09
MMLU-Pro87.8%Benchgen evaluation2025-09
SWE-Bench Verified76.4%Benchgen evaluation2025-09

Qwen3.5 397B A17B vs Alternatives

ModelBrowseCompHLEMMLU-ProLicenseActive params
Qwen3.5 397B A17B69.0%28.7%87.8%Apache 2.017B
Qwen3.5 122B A10B63.8%47.5%86.7%Apache 2.010B
Kimi K2.683.2%36.4%Apache 2.0MoE

Qwen3.5 397B A17B vs 122B A10B: higher BrowseComp (69% vs 63.8%) but lower HLE (28.7% vs 47.5%) despite being larger. Use 122B A10B for reasoning/HLE tasks; 397B A17B for BrowseComp-heavy web research.

Frequently Asked Questions

What is Qwen3.5 397B A17B? Alibaba's September 2025 large MoE model: 397B total / 17B active parameters, scoring 69.0% BrowseComp, 88.4% GPQA Diamond, 87.8% MMLU-Pro. Apache 2.0.

Specs from Alibaba's Qwen3.5 397B A17B release (September 2025) and Benchgen evaluations. Last updated 2026-07-24.

Benchmark Leaderboards

This model isn’t on any benchmark leaderboard yet.