Benchgen
Models/meta/

Llama 3.2 3B Instruct

DraftPublic

Model Details

Llama 3.2 3B Instruct

Organization Parameters Context License Weights Released

Quick answer: Llama 3.2 3B Instruct is Meta's September 2024 small open-weight model, scoring 77.7% on GSM8K, 69.8% on HellaSwag, and 67.0% on BFCL v2. With a 128K context window and Llama 3.2 License, it is designed for on-device deployment and edge inference.

At a Glance

Where Llama 3.2 3B leads

  • 77.7% GSM8K — strong math reasoning at 3B scale
  • 128K context window — unusually long for a 3B model
  • On-device capable — fits mobile and edge hardware
  • 67.0% BFCL v2 — solid tool use at 3B scale
  • Llama 3.2 License permits on-device distribution
  • Strong foundation for 3B-scale fine-tuning

Where it lags

  • 69.8% HellaSwag — common-sense reasoning below larger models
  • Text-only in the 3B variant (Llama 3.2 11B/90B add vision)
  • Significantly below 8B+ models on all benchmarks
  • December 2023 knowledge cutoff (older than 70B variant's March 2024)

Best for: On-device inference; mobile applications; edge deployments; rapid response applications requiring low latency at 3B scale.

What Llama 3.2 3B Instruct Is

Llama 3.2 3B Instruct (released September 25, 2024) is Meta's smallest instruction-tuned model in the Llama 3.2 family. Its primary design goal is on-device and edge deployment — running on mobile phones, laptops, and embedded hardware without cloud inference.

Despite its small size, Llama 3.2 3B achieves 77.7% GSM8K — a score that would have been competitive with state-of-the-art models just two years earlier. Its 128K context window is exceptionally long for a 3B model, enabling long-document tasks at the edge.

The Llama 3.2 3B is part of Meta's broader push for on-device AI with the Llama 3.2 family. The 1B variant goes even smaller for more constrained hardware; the 11B and 90B variants in the same family add multimodal vision capability.

Specifications

FieldValue
OrganizationMeta
Parameters3B (dense)
Context window128,000 tokens
LicenseLlama 3.2 License
HuggingFacemeta-llama/Llama-3.2-3B-Instruct
Release dateSeptember 25, 2024
Knowledge cutoffDecember 2023
ModalityText only

Pricing

Llama 3.2 3B Instruct is available as open weights on Hugging Face under the Llama 3.2 License. Available via numerous hosted providers at extremely low cost due to small model size.

Public Benchmark Scores

BenchmarkScoreSourceDate
GSM8K77.7%Benchgen evaluation2025-07
HellaSwag69.8%Benchgen evaluation2025-07
BFCL v267.0%Benchgen evaluation2025-07

Llama 3.2 3B vs Alternatives

ModelGSM8KHellaSwagBFCL v2Params
Llama 3.2 3B Instruct77.7%69.8%67.0%3B
Llama 3.1 8B Instruct84.5%8B
Phi-414B

Llama 3.2 3B is designed for on-device use where the 8B model won't fit. For server-side deployments, Llama 3.1 8B Instruct provides meaningfully better scores (84.5% vs 77.7% GSM8K) at modest cost increase.

Frequently Asked Questions

What is Llama 3.2 3B Instruct? Llama 3.2 3B Instruct is Meta's September 2024 small on-device model scoring 77.7% GSM8K, 69.8% HellaSwag, and 67.0% BFCL v2 with a 128K context window.
Can Llama 3.2 3B run on-device? Yes — that is its primary design goal. At 3B parameters (quantised to 4-bit, ~1.8GB), it fits on most modern smartphones and edge hardware.
Should I use Llama 3.2 3B or Llama 3.1 8B? Use 3B for on-device or edge deployment where 8B won't fit. For server-side inference, Llama 3.1 8B Instruct provides significantly better benchmark performance (84.5% vs 77.7% GSM8K) at low cost.

Specs from Meta's official Llama 3.2 release (September 2024) and Benchgen evaluations. Last updated 2026-07-24.

Benchmark Leaderboards

This model isn’t on any benchmark leaderboard yet.