Quick answer: Gemma 4 31B is the flagship dense open-weight model in Google DeepMind's Gemma 4 family and the successor to Gemma 3 27B. With 31B parameters, it delivers Gemma 4-generation training improvements — better coding, stronger multilingual capability, and improved instruction following — at the cost of a single 2× A100 or H100 80GB setup.
Gemma 4 31B is Google DeepMind's largest dense open-weight model in the Gemma 4 generation. Building on the Gemma 3 27B architecture, the slight parameter increase (27B → 31B) reflects architecture refinements for better utilisation efficiency, while the training improvements from Gemma 4's development cycle deliver meaningful quality gains on reasoning, coding, and multilingual tasks.
For Benchgen evaluation purposes, Gemma 4 31B is the primary open-weight model to benchmark against the 30–35B closed-API tier (GPT-5.1 Instant, Claude Haiku 4.5).
| Field | Value |
|---|---|
| Organization | Google DeepMind |
| Parameters | 31 billion |
| License | Gemma Terms of Use |
| Modality | Multimodal (text and vision) |
| Min VRAM (FP16) | ~62GB (2× A100 40GB or 1× H100 80GB) |
| HuggingFace | google/gemma-4-31b-it |
Last updated 2026-06-19.
Open weights — self-host at no cost.
| Model | MMLU-Pro | License |
|---|---|---|
| Gemma 4 31B | 85.2% | Gemma ToU |
| Gemma 4 26B A4B | — | — |
Check the MMLU-Pro leaderboard for current comparisons.
Scores from Benchgen evaluations. Last updated 2026-07-24.