Quick answer: Qwen3.5 122B A10B is Alibaba's September 2025 MoE model scoring 63.8% BrowseComp, 47.5% HLE, 86.6% GPQA Diamond, 86.7% MMLU-Pro, and 72% SWE-Bench Verified. Apache 2.0 with 122B total / 10B active parameters.
Where Qwen3.5 122B A10B leads
Where it lags
Best for: Open-weight frontier reasoning with 47.5% HLE; web research at 63.8% BrowseComp; academic knowledge tasks; teams needing Apache 2.0 models with Qwen 3.5 generation capabilities.
Qwen3.5 122B A10B is the September 2025 flagship in Alibaba's Qwen3.5 generation — succeeding Qwen3 (April/May 2025). As a 122B/10B MoE, inference activates only 10B parameters per token while the full 122B capacity is available for specialised knowledge routing.
The 47.5% HLE score is notably high for an open-weight model — near Gemini 3.1 Pro (46.4%) and GLM-5.2 (54.7%) while being Apache 2.0 open. This makes it one of the strongest openly-available models for academic and research reasoning.
| Field | Value |
|---|---|
| Organization | Alibaba |
| License | Apache 2.0 |
| HuggingFace | Qwen/Qwen3.5-122B-A10B |
| Release date | September 2025 |
| Parameters | 122B total / 10B active (MoE) |
| Modality | Text only |
| Context window | 128K tokens |
Open weights under Apache 2.0 — self-host. Also via Alibaba Cloud DashScope.
| Benchmark | Score | Source | Date |
|---|---|---|---|
| BrowseComp | 63.8% | Benchgen evaluation | 2025-09 |
| GPQA Diamond | 86.6% | Benchgen evaluation | 2025-09 |
| Humanity's Last Exam | 47.5% | Benchgen evaluation | 2025-09 |
| MMLU-Pro | 86.7% | Benchgen evaluation | 2025-09 |
| SWE-Bench Verified | 72% | Benchgen evaluation | 2025-09 |
| Model | HLE | BrowseComp | MMLU-Pro | License |
|---|---|---|---|---|
| Qwen3.5 122B A10B | 47.5% | 63.8% | 86.7% | Apache 2.0 |
| Kimi K2 Thinking 0905 | 51.0% | 60.2% | 84.6% | Apache 2.0 |
| Kimi K2.6 | 36.4% | 83.2% | — | Apache 2.0 |
| Gemini 3 Flash | 43.5% | — | — | Proprietary |
Qwen3.5 122B A10B vs Kimi K2 Thinking 0905: higher HLE (47.5% vs 51%), higher MMLU-Pro (86.7% vs 84.6%), higher BrowseComp (63.8% vs 60.2%). Very close competition; both Apache 2.0.
Specs from Alibaba's Qwen3.5 122B A10B release (September 2025) and Benchgen evaluations. Last updated 2026-07-24.
This model isn’t on any benchmark leaderboard yet.