Benchgen
Models/meta/

LlamaGuard-4-12B

DraftPublic

Model Details

LlamaGuard-4-12B

Organization Pricing License Modality

Quick answer: LlamaGuard-4-12B is Meta's 12B multimodal safety classifier, the successor to the LlamaGuard line, supporting both text and image moderation. It's the largest model in Shieldstral's comparison set that reports scores across all task categories including multimodal, but generally trails Shieldstral, GPT-OSS-Safeguard, and Qwen3Guard on text tasks.

At a Glance

Where it leads: Not the top performer on any individual benchmark in Shieldstral's comparison, but the only large (12B) multimodal model that reports scores across the full range of text and image tasks. Where it lags: WildGuardTest (Prompt/Response), OpenAI Moderation, and BeaverTails, where smaller Shieldstral and Qwen3Guard-8B outperform it. Best for: Teams already in the Llama ecosystem needing a single multimodal guard model with broad task coverage.

What LlamaGuard-4-12B Is

LlamaGuard-4-12B builds on Meta's Llama Guard series (LlamaGuard, LlamaGuard 2, LlamaGuard 3), extending native multimodal support to the 12B parameter class. Like its predecessors, it classifies content against a defined taxonomy of unsafe categories for both prompts and responses, now extended to images.

Specifications

FieldValue
OrganizationMeta
Parameters12B
LicenseLlama 4 Community License
ModalityMultimodal (text + image)

Pricing

Open weights, free to download under the Llama 4 Community License (usage restrictions apply for very large deployments).

Public Benchmark Scores

Scores from Mistral AI's Shieldstral model card, shown for context. Not Benchgen measurements.

Frequently Asked Questions

What is LlamaGuard-4-12B?Meta's 12B multimodal open-weights safety classifier, the latest in the Llama Guard series, supporting both text and image moderation.
Is it multimodal?Yes, it natively supports image inputs alongside text.
Is it open source?Yes, released under the Llama 4 Community License.

Scores sourced from Mistral AI's Shieldstral model card, shown for context. Last updated 2026-08-12.

Benchmark Leaderboards

This model isn’t on any benchmark leaderboard yet.