Quick answer: Qwen3 VL 30B A3B Instruct is Alibaba's February 2026 compact MoE vision-language instruct model scoring 77.8% MMLU-Pro, 66.3% BFCL-v3, and 58.5% Arena Hard v2. Apache 2.0 — 3B active params.
Where Qwen3 VL 30B A3B Instruct leads
Where it lags
Best for: Cost-efficient VL instruct deployments; faster inference than thinking variant.
The standard instruct (non-thinking) 30B A3B MoE variant. Faster inference than the thinking variant, with slightly lower quality.
| Field | Value |
|---|---|
| Organization | Alibaba |
| License | Apache 2.0 |
| HuggingFace | Qwen/Qwen3-VL-30B-A3B-Instruct |
| Release date | February 2026 |
| Parameters | 30B total / 3B active (MoE) |
| Modality | Text and vision |
Open weights under Apache 2.0 — self-host at no cost.
| Benchmark | Score | Source | Date |
|---|---|---|---|
| MMLU-Pro | 77.8% | Benchgen evaluation | 2026-02 |
| BFCL-v3 | 66.3% | Benchgen evaluation | 2026-02 |
| Arena Hard v2 | 58.5% | Benchgen evaluation | 2026-02 |
| Model | MMLU-Pro | BFCL-v3 | Arena Hard | Type |
|---|---|---|---|---|
| Qwen3 VL 30B A3B Instruct | 77.8% | 66.3% | 58.5% | MoE Instruct |
| Qwen3 VL 30B A3B Thinking | 80.5% | 68.6% | 56.7% | MoE Thinking |
| Qwen3 VL 32B Instruct | 78.6% | 70.2% | 64.7% | Dense Instruct |
For maximum Arena Hard: 32B Instruct (64.7%). For MMLU-Pro at 3B active: 30B A3B Thinking (80.5%).
Specs from Alibaba's Qwen3 VL 30B A3B Instruct release (February 2026) and Benchgen evaluations. Last updated 2026-07-24.
This model isn’t on any benchmark leaderboard yet.