Benchgen
Models/mistral/

Mistral Large 2

DraftPublic

Model Details

Mistral Large 2

Organization Context License Released

Quick answer: Mistral Large 2 is Mistral's July 2024 flagship model, scoring 93.0% on GSM8K, 92.0% on HumanEval, 84.0% on MMLU, and 8.63/10 on MT-Bench. With a 128K context window, it was Mistral's most capable model at launch and remains relevant for teams in the Mistral AI ecosystem.

At a Glance

Where Mistral Large 2 leads

  • 92.0% HumanEval — excellent code generation
  • 93.0% GSM8K — strong math reasoning
  • 8.63 MT-Bench — high-quality instruction following
  • 128K context window
  • Available on HuggingFace (Mistral Research License)
  • Strong function calling and tool use

Where it lags

  • Mistral Research License — no commercial deployment without Mistral AI agreement
  • Text-only; no vision
  • 84.0% MMLU — below frontier models on academic knowledge
  • Superseded by Mistral Large 2407 and Codestral for specific tasks

Best for: Research and evaluation within the Mistral AI ecosystem; teams with Mistral Enterprise agreements; code generation tasks.

What Mistral Large 2 Is

Mistral Large 2 (released July 24, 2024) was Mistral AI's highest-capability model in their "Large" tier, positioned above Mistral Large (2402) and below proprietary enterprise models. It offers strong coding (92.0% HumanEval) and reasoning (93.0% GSM8K) at 128K context.

The model is available on HuggingFace under the Mistral Research License — meaning the weights are accessible but commercial deployment requires a Mistral AI Enterprise agreement. For teams comparing open-weight options, Llama 3.3 70B (Llama 3.3 License) or DeepSeek-V3 (MIT) offer similar or better benchmarks with more permissive commercial terms.

Specifications

FieldValue
OrganizationMistral AI
Context window128,000 tokens
LicenseMistral Research License
HuggingFacemistralai/Mistral-Large-Instruct-2407
Release dateJuly 24, 2024
Knowledge cutoffJune 2024
ModalityText only

Pricing

Available via Mistral AI API (La Plateforme) with commercial pricing. Research use via HuggingFace under Mistral Research License.

Public Benchmark Scores

BenchmarkScoreSourceDate
GSM8K93.0%Benchgen evaluation2025-07
HumanEval92.0%Benchgen evaluation2025-07
MMLU84.0%Benchgen evaluation2025-07
MT-Bench8.63 / 10Benchgen evaluation2025-07

Mistral Large 2 vs Alternatives

ModelGSM8KHumanEvalMMLULicense
Mistral Large 293.0%92.0%84.0%Research only
Llama 3.3 70B Instruct88.4%86.0%Llama 3.3
DeepSeek-V388.5%MIT

Mistral Large 2's coding scores (92% HumanEval) are competitive with Llama 3.3 70B (88.4%), but the Research License restricts commercial use. For permissive commercial deployment, Llama 3.3 70B or DeepSeek-V3 are preferred.

Frequently Asked Questions

What is Mistral Large 2? Mistral Large 2 is Mistral AI's July 2024 flagship model scoring 93.0% GSM8K, 92.0% HumanEval, 84.0% MMLU, and 8.63 MT-Bench with 128K context.
Is Mistral Large 2 open source? The weights are available under the Mistral Research License — free for research, but commercial deployment requires a Mistral AI Enterprise agreement.

Specs from Mistral AI's official Mistral Large 2 release (July 2024) and Benchgen evaluations. Last updated 2026-07-24.

Benchmark Leaderboards

This model isn’t on any benchmark leaderboard yet.