Benchgen
Models/mistral-ai/

Mistral Small 3.1 24B Instruct

DraftPublic

Model Details

Mistral Small 3.1 24B Instruct

Organization License Released

Quick answer: Mistral Small 3.1 24B Instruct is Mistral AI's March 2025 multimodal update scoring 88.4% HumanEval, 69.3% MATH, 66.8% MMLU-Pro, and 10.4% SimpleQA. Apache 2.0.

At a Glance

Where Mistral Small 3.1 leads

  • 88.4% HumanEval — excellent coding at 24B
  • 69.3% MATH — solid competition math
  • Apache 2.0 — fully open
  • Multimodal: vision support added

Where it lags

  • 10.4% SimpleQA — low factual accuracy
  • 66.8% MMLU-Pro — moderate advanced knowledge

Best for: Apache 2.0 coding (HumanEval) with multimodal support; Small 3 series users needing vision capability.

What Mistral Small 3.1 24B Is

Mistral Small 3.1 24B Instruct is the March 2025 update to Mistral Small 3, adding vision/multimodal support to the 24B model. The 88.4% HumanEval is a notable score — competitive with models twice its size.

Specifications

FieldValue
OrganizationMistral AI
LicenseApache 2.0
HuggingFacemistralai/Mistral-Small-3.1-24B-Instruct
Release dateMarch 17, 2025
Parameters24B
ModalityText and vision (multimodal)
Context window128K tokens

Pricing

Open weights under Apache 2.0. Also available via Mistral API.

Public Benchmark Scores

BenchmarkScoreSourceDate
HumanEval88.4%Benchgen evaluation2025-03
MATH69.3%Benchgen evaluation2025-03
MMLU-Pro66.8%Benchgen evaluation2025-03
SimpleQA10.4%Benchgen evaluation2025-03

Mistral Small 3.1 vs Alternatives

ModelHumanEvalMATHMMLU-ProMultimodal
Mistral Small 3.1 24B88.4%69.3%66.8%Yes
Mistral Small 3 24B84.8%70.6%66.3%No
Mistral Small 3.2 24B69.4%69.1%Yes

Mistral Small 3.1 vs 3.0: higher HumanEval (88.4% vs 84.8%), similar MATH. 3.1 adds multimodal. Choose 3.0 for Arena Hard instruction tasks.

Frequently Asked Questions

What is Mistral Small 3.1 24B Instruct? Mistral AI's March 2025 multimodal 24B model scoring 88.4% HumanEval, 69.3% MATH. Apache 2.0 — adds vision support to Small 3.

Specs from Mistral AI's Small 3.1 24B Instruct release (March 2025) and Benchgen evaluations. Last updated 2026-07-24.