Benchgen
Models/moonshot-ai/

Kimi-VL-A3B-Thinking-2506

DraftPublic

Model Details

Kimi-VL-A3B-Thinking-2506

Organization Modality License Released

Quick answer: Kimi-VL-A3B-Thinking-2506 is Moonshot AI's mixture-of-experts multimodal reasoning model (2.8B activated parameters), which ranked second on the Video-MMMU knowledge-acquisition leaderboard among all evaluated models.

At a Glance

Where Kimi-VL-A3B-Thinking-2506 leads

  • Strong efficiency-to-performance ratio via a sparse mixture-of-experts architecture with only ~3B activated parameters
  • Ranked second overall on the Video-MMMU knowledge-acquisition benchmark

Where it lags

  • Smaller activated parameter count than dense frontier models can limit some general reasoning tasks

Best for: teams needing efficient, strong multimodal reasoning performance from a compact mixture-of-experts model.

What Kimi-VL-A3B-Thinking-2506 Is

Kimi-VL-A3B-Thinking-2506 is a mixture-of-experts multimodal reasoning model from Moonshot AI, part of the Kimi-VL family, using a "thinking" (extended chain-of-thought) mode to improve reasoning quality. It ranked second overall on the Video-MMMU benchmark, which tests knowledge acquisition from expert-level lecture videos.

Specifications

FieldValue
OrganizationMoonshot AI
ModalityMultimodal (text + vision)
LicenseOpen weights (MIT)
Release dateJune 2025