Quick answer: Gemma 2 9B is Google's July 2024 mid-size open-weight model, scoring 88.0% on ARC-E, 68.6% on GSM8K, and 81.9% on HellaSwag. It was designed for efficient deployment at 9B scale, but has since been superseded by Gemma 3 12B with a much larger 128K context window.
Where Gemma 2 9B leads
Where it lags
Best for: Existing Gemma 2 9B integrations; legacy support.
Gemma 2 9B (released July 31, 2024) was Google's mid-tier Gemma 2 model, positioned between the 2B and 27B variants. Like Gemma 2 27B, its primary limitation is the 8K context window.
For new deployments at the 9-12B scale, Gemma 3 12B is the recommended replacement: 128K context, multimodal vision capability, better benchmarks, and same Gemma Terms of Use license. Llama 3.1 8B (Apache 2.0, 128K context) is also preferred if a fully permissive license is needed.
| Field | Value |
|---|---|
| Organization | |
| Parameters | 9B (dense) |
| Context window | 8,192 tokens |
| License | Gemma Terms of Use |
| HuggingFace | google/gemma-2-9b-it |
| Release date | July 31, 2024 |
| Knowledge cutoff | June 2024 |
| Modality | Text only |
Open weights under Gemma Terms of Use — free to self-host.
| Benchmark | Score | Source | Date |
|---|---|---|---|
| ARC-E | 88.0% | Benchgen evaluation | 2025-07 |
| GSM8K | 68.6% | Benchgen evaluation | 2025-07 |
| HellaSwag | 81.9% | Benchgen evaluation | 2025-07 |
| Model | Context | GSM8K | Vision | License |
|---|---|---|---|---|
| Gemma 2 9B | 8K | 68.6% | No | Gemma ToU |
| Gemma 3 12B | 128K | — | Yes | Gemma ToU |
| Llama 3.1 8B Instruct | 128K | 84.5% | No | Apache 2.0 |
For new deployments: Gemma 3 12B (128K, vision, same license) or Llama 3.1 8B (Apache 2.0, 128K, higher GSM8K).
Specs from Google's official Gemma 2 release (July 2024) and Benchgen evaluations. Last updated 2026-07-24.
This model isn’t on any benchmark leaderboard yet.