Benchgen
Models/alibaba/

Qwen3 VL 30B A3B Thinking

DraftPublic

Model Details

Qwen3 VL 30B A3B Thinking

Organization License Released MoE

Quick answer: Qwen3 VL 30B A3B Thinking is Alibaba's February 2026 compact MoE vision-language thinking model scoring 80.5% MMLU-Pro, 68.6% BFCL-v3, and 56.7% Arena Hard v2. Apache 2.0 — 3B active params.

At a Glance

Where Qwen3 VL 30B A3B Thinking leads

  • 80.5% MMLU-Pro — nearly matches 32B dense Thinking (82.1%)
  • 3B active params: low inference cost
  • Apache 2.0 — fully open

Where it lags

  • 56.7% Arena Hard v2 — below 32B Thinking (60.5%) and 32B Instruct (64.7%)

Best for: Cost-efficient VL thinking at 3B active params; MMLU-Pro heavy vision tasks.

What Qwen3 VL 30B A3B Thinking Is

A 30B MoE (3B active params) thinking variant for vision-language tasks. Near-equal MMLU-Pro to the 32B dense thinking at a fraction of the inference cost.

Specifications

FieldValue
OrganizationAlibaba
LicenseApache 2.0
HuggingFaceQwen/Qwen3-VL-30B-A3B-Thinking
Release dateFebruary 2026
Parameters30B total / 3B active (MoE)
ModalityText and vision

Pricing

Open weights under Apache 2.0 — self-host at no cost.

Public Benchmark Scores

BenchmarkScoreSourceDate
MMLU-Pro80.5%Benchgen evaluation2026-02
BFCL-v368.6%Benchgen evaluation2026-02
Arena Hard v256.7%Benchgen evaluation2026-02

Qwen3 VL 30B A3B Thinking vs Alternatives

ModelMMLU-ProBFCL-v3Arena HardActive Params
Qwen3 VL 30B A3B Thinking80.5%68.6%56.7%3B
Qwen3 VL 32B Thinking82.1%71.7%60.5%32B
Qwen3 VL 30B A3B Instruct77.8%66.3%58.5%3B

30B A3B Thinking vs 32B Thinking: nearly same MMLU-Pro (80.5% vs 82.1%) at 1/10 the active params. Best cost/quality in Qwen3 VL thinking family.

Frequently Asked Questions

What is Qwen3 VL 30B A3B Thinking? Alibaba's February 2026 compact MoE VL thinking model (30B/3B active) scoring 80.5% MMLU-Pro. Apache 2.0 — best cost/quality in Qwen3 VL thinking family.

Specs from Alibaba's Qwen3 VL 30B A3B Thinking release (February 2026) and Benchgen evaluations. Last updated 2026-07-24.

Benchmark Leaderboards

This model isn’t on any benchmark leaderboard yet.