Benchgen
Models/alibaba/

Qwen3 VL 32B Instruct

DraftPublic

Model Details

Qwen3 VL 32B Instruct

Organization License Released VL

Quick answer: Qwen3 VL 32B Instruct is Alibaba's February 2026 32B vision-language instruct model scoring 78.6% MMLU-Pro, 70.2% BFCL-v3, and 64.7% Arena Hard v2. Apache 2.0.

At a Glance

Where Qwen3 VL 32B Instruct leads

  • 70.2% BFCL-v3 — strong function calling
  • 64.7% Arena Hard v2 — better instruction than 30B/3B variant
  • Apache 2.0 — fully open

Where it lags

  • Below Qwen3 VL 32B Thinking on MMLU-Pro (78.6% vs 82.1%)
  • Dense 32B: higher inference cost than 30B A3B MoE

Best for: Vision-language function calling with 32B quality; non-thinking instruct deployments in Qwen3 VL.

What Qwen3 VL 32B Instruct Is

Qwen3 VL 32B Instruct is the standard instruct (non-thinking) 32B vision-language model in the Qwen3 VL family. Compared to the 32B Thinking variant, it scores lower on MMLU-Pro (78.6% vs 82.1%) but is faster to run (no long chain-of-thought).

Specifications

FieldValue
OrganizationAlibaba
LicenseApache 2.0
HuggingFaceQwen/Qwen3-VL-32B-Instruct
Release dateFebruary 2026
Parameters32B (dense)
ModalityText and vision

Pricing

Open weights under Apache 2.0 — self-host at no cost.

Public Benchmark Scores

BenchmarkScoreSourceDate
MMLU-Pro78.6%Benchgen evaluation2026-02
BFCL-v370.2%Benchgen evaluation2026-02
Arena Hard v264.7%Benchgen evaluation2026-02

Qwen3 VL 32B Instruct vs Alternatives

ModelMMLU-ProBFCL-v3Arena HardType
Qwen3 VL 32B Instruct78.6%70.2%64.7%Instruct
Qwen3 VL 32B Thinking82.1%71.7%60.5%Thinking
Qwen3 VL 30B A3B Instruct77.8%66.3%58.5%MoE Instruct

32B Instruct vs 32B Thinking: Instruct has higher Arena Hard (64.7% vs 60.5%) but lower MMLU-Pro (78.6% vs 82.1%). Use Thinking for knowledge/reasoning; Instruct for instruction following speed.

Frequently Asked Questions

What is Qwen3 VL 32B Instruct? Alibaba's February 2026 vision-language instruct model scoring 78.6% MMLU-Pro, 64.7% Arena Hard v2. Apache 2.0.

Specs from Alibaba's Qwen3 VL 32B Instruct release (February 2026) and Benchgen evaluations. Last updated 2026-07-24.

Benchmark Leaderboards

This model isn’t on any benchmark leaderboard yet.