Benchgen
Models/deepseek/

DeepSeek-V3.2-Speciale

DraftPublic

Model Details

DeepSeek-V3.2 Speciale

Organization License Weights Released

Quick answer: DeepSeek-V3.2 Speciale is a specialized October 2025 variant in the DeepSeek-V3.2 family, scoring 73.1% on SWE-Bench Verified and 30.6% on Humanity's Last Exam. Under MIT with open weights, it is specifically optimised for software engineering tasks — the "Speciale" name indicating targeted task specialisation.

At a Glance

Where DeepSeek-V3.2 Speciale leads

  • 73.1% SWE-Bench Verified — state-of-the-art software engineering among open models at launch
  • MIT license — fully permissive commercial use
  • Open weights for self-hosting
  • Specialised SWE training vs base V3.2 Experimental

Where it lags

  • 30.6% Humanity's Last Exam — weaker on frontier academic tasks (vs 97.1% SimpleQA of V3.2 Exp)
  • Text-only
  • Limited public benchmark data vs V3.2 Experimental

Best for: Open-weight SWE-Bench class software engineering; agentic code-fixing pipelines; MIT-licensed self-hosted coding agents.

What DeepSeek-V3.2 Speciale Is

DeepSeek-V3.2 Speciale is a task-specialized variant of V3.2, fine-tuned for software engineering (SWE-Bench) performance. The "Speciale" designation (Italian for "special/specialized") indicates targeted training on code debugging, patching, and repository-level software engineering tasks.

Its 73.1% SWE-Bench Verified score was among the highest for an open-weight model at launch, demonstrating that specialised fine-tuning on SWE tasks can close much of the gap with larger proprietary models on this benchmark.

The MIT license and open weights make Speciale the preferred choice for teams building open-source coding agents or agentic software engineering pipelines where self-hosting is required.

Specifications

FieldValue
OrganizationDeepSeek
LicenseMIT
HuggingFacedeepseek-ai/DeepSeek-V3-2-Speciale
Release dateOctober 2025
Knowledge cutoffJune 2025
ModalityText only

Pricing

Open weights under MIT — free to self-host. Available via DeepSeek API.

Public Benchmark Scores

BenchmarkScoreSourceDate
SWE-Bench Verified73.1%Benchgen evaluation2025-10
Humanity's Last Exam30.6%Benchgen evaluation2025-10

DeepSeek-V3.2 Speciale vs Alternatives

ModelSWE-Bench VerifiedLicenseSpecialisation
DeepSeek-V3.2 Speciale73.1%MITSWE-focused
DeepSeek-V3.2 ExperimentalMITGeneral
DeepSeek-V4 Flash Max79%ProprietaryGeneral

V3.2 Speciale vs V4 Flash Max: lower SWE-Bench (73.1% vs 79%) but MIT vs Proprietary. For open-weight SWE pipelines: Speciale. For maximum SWE performance: V4 Flash Max.

Frequently Asked Questions

What is DeepSeek-V3.2 Speciale? DeepSeek's October 2025 SWE-specialised MIT model scoring 73.1% SWE-Bench Verified and 30.6% HLE. Open weights, optimised for software engineering agent tasks.
What is the difference between V3.2 Speciale and V3.2 Experimental? V3.2 Experimental (97.1% SimpleQA, 74.1% LiveCodeBench) is a general-purpose improvement. V3.2 Speciale is specialised for SWE-Bench-class software engineering. Both are MIT under open weights.

Specs from DeepSeek's V3.2 Speciale release (October 2025) and Benchgen evaluations. Last updated 2026-07-24.

Benchmark Leaderboards

This model isn’t on any benchmark leaderboard yet.