Benchgen
Models/tu-darmstadt-hessian-ai/

LlavaGuard-7B

DraftPublic

Model Details

LlavaGuard-7B

Organization Pricing License Modality

Quick answer: LlavaGuard-7B is TU Darmstadt and hessian.AI's 7B open-weights vision-language safety classifier, judging image content against a detailed, structured safety policy rather than a fixed label set — the model behind the LlavaGuard benchmark. It's the strongest performer on its own LlavaGuard benchmark and competitive on VLGuard and UnsafeBench in Mistral's Shieldstral comparison.

At a Glance

Where it leads: LlavaGuard — naturally the top scorer on its own namesake benchmark and evaluation methodology. Where it lags: VLGuard and UnsafeBench, where Shieldstral's smaller, more recent multimodal design scores higher. Best for: Teams needing an academic-grade, policy-conditioned image moderation classifier with a research paper trail.

What LlavaGuard-7B Is

LlavaGuard-7B is built on the LLaVA vision-language architecture, fine-tuned to judge image content against detailed written safety policies (category, rationale, and safety rating) rather than a fixed category label — an approach conceptually similar to Shieldstral's policy-adaptive design but purpose-built for image moderation specifically.

Specifications

FieldValue
OrganizationTU Darmstadt / hessian.AI
Parameters7B
Base architectureLLaVA
LicenseApache 2.0
ModalityMultimodal (image + policy text)

Pricing

Open weights, free to download.

Public Benchmark Scores

Scores from Mistral AI's Shieldstral model card, shown for context. Not Benchgen measurements.

BenchmarkScore
LlavaGuard81.4%
VLGuard69.5%
UnsafeBench63.9%

Frequently Asked Questions

What is LlavaGuard-7B?TU Darmstadt and hessian.AI's 7B open-weights vision-language safety classifier that judges image content against detailed written safety policies.
Is it multimodal?Yes, it's built on the LLaVA vision-language architecture.
Is it open source?Yes, released under Apache 2.0.

Scores sourced from Mistral AI's Shieldstral model card, shown for context. Last updated 2026-08-12.

Benchmark Leaderboards

This model isn’t on any benchmark leaderboard yet.