Benchgen
Models/google/

ShieldGemma-2-4B

DraftPublic

Model Details

ShieldGemma-2-4B

Organization Pricing License Modality

Quick answer: ShieldGemma 2 (4B) is Google's 4B multimodal safety classifier built on Gemma 3, focused primarily on image content moderation (unsafe image categories like dangerous content, sexually explicit material, and violence). It's Google's successor to the text-only ShieldGemma-9B, adding native vision support in a smaller footprint.

At a Glance

Where it leads: Not a top scorer in Shieldstral's published comparison set (limited public score overlap). Where it lags: No text-only prompt/response classification scores reported in Mistral's comparison — ShieldGemma 2 focuses on image classification specifically. Best for: Teams needing a lightweight, Gemma-family image safety classifier alongside a separate text moderation stack.

What ShieldGemma-2-4B Is

ShieldGemma 2 is built on Gemma 3's 4B vision-language model, specialized for classifying images as safe or unsafe across a defined harm taxonomy. It represents Google's shift toward native multimodal safety classification, succeeding the text-only ShieldGemma-9B.

Specifications

FieldValue
OrganizationGoogle
Parameters4B
Base modelGemma 3
LicenseGemma Terms of Use
ModalityMultimodal (primarily image)

Pricing

Open weights, free to download under Gemma's Terms of Use.

Public Benchmark Scores

Scores from Mistral AI's Shieldstral model card, shown for context. Not Benchgen measurements.

BenchmarkScore
VLGuard61.3%
UnsafeBench54.9%
LlavaGuard56.2%

Frequently Asked Questions

What is ShieldGemma 2 (4B)?Google's 4B multimodal safety classifier built on Gemma 3, focused on image content moderation.
Is it multimodal?Yes, it natively classifies images; it is Google's vision-focused successor to the text-only ShieldGemma-9B.
Is it open source?Yes, released under Google's Gemma Terms of Use.

Last updated 2026-08-12.

Benchmark Leaderboards

This model isn’t on any benchmark leaderboard yet.