Benchgen
Models/mistral-ai/

Mistral Small 3.2 24B Instruct

DraftPublic

Model Details

Mistral Small 3.2 24B Instruct

Organization License Released

Quick answer: Mistral Small 3.2 24B Instruct is Mistral AI's June 2025 update scoring 43.1% Arena Hard, 69.4% MATH, 69.1% MMLU-Pro, and 12.1% SimpleQA. Apache 2.0.

At a Glance

Where Mistral Small 3.2 leads

  • 69.4% MATH — solid competition math
  • 69.1% MMLU-Pro — reasonable academic breadth
  • Apache 2.0 — fully open

Where it lags

  • 43.1% Arena Hard — significantly below Mistral Small 3 (87.6%) — regression
  • 12.1% SimpleQA — low factual accuracy
  • Below Mistral Small 3.1 on Arena Hard (87.6% vs 43.1%)

Best for: Math-intensive pipelines needing Apache 2.0; use Mistral Small 3 for instruction quality; prefer 3.1 for Arena Hard tasks.

What Mistral Small 3.2 24B Is

Mistral Small 3.2 24B Instruct is Mistral AI's June 2025 update to the Small 3 series. Notably, the Arena Hard drops significantly (43.1% vs 87.6% in Small 3) — suggesting this version trades instruction quality for other improvements.

For instruction-following tasks: use Mistral Small 3 (87.6% Arena Hard) or Mistral Small 3.1 (88.4% HumanEval). For math: 3.2 is comparable at 69.4%.

Specifications

FieldValue
OrganizationMistral AI
LicenseApache 2.0
HuggingFacemistralai/Mistral-Small-3.2-24B-Instruct
Release dateJune 2025
Parameters24B
ModalityText and vision

Pricing

Open weights under Apache 2.0. Also available via Mistral API.

Public Benchmark Scores

BenchmarkScoreSourceDate
Arena Hard43.1%Benchgen evaluation2025-06
MATH69.4%Benchgen evaluation2025-06
MMLU-Pro69.1%Benchgen evaluation2025-06
SimpleQA12.1%Benchgen evaluation2025-06

Mistral Small 3.2 vs Alternatives

ModelArena HardMATHMMLU-ProLicense
Mistral Small 3.2 24B43.1%69.4%69.1%Apache 2.0
Mistral Small 3 24B87.6%70.6%66.3%Apache 2.0
Mistral Small 3.1 24B69.3%66.8%Apache 2.0

Mistral Small 3.2 has lower Arena Hard than 3.0 (43.1% vs 87.6%) — use Small 3 for instruction quality; use 3.2 only for multimodal or specific math tasks.

Frequently Asked Questions

What is Mistral Small 3.2 24B Instruct? Mistral AI's June 2025 24B model scoring 43.1% Arena Hard, 69.4% MATH, 69.1% MMLU-Pro. Apache 2.0 — lower Arena Hard than Small 3; use Small 3 for instruction tasks.

Specs from Mistral AI's Small 3.2 24B Instruct release (June 2025) and Benchgen evaluations. Last updated 2026-07-24.