Skip to content

Llama Guard 4 12B

Also accepted:meta-llama/llama-guard-4-12b

Llama Guard 4 is a 12B natively multimodal safety classifier for content safety classification on LLM inputs and responses. It acts as an LLM itself: it generates text indicating whether a prompt or response is safe or unsafe and, if unsafe, lists the violated categories (MLCommons hazards taxonomy S1–S14). Use it as a guardrail/judge hop in front of or behind chat models.

Providers
Capabilities
DeepInfra
deepinfra-byok
$0
$0
Unavailable
OpenRouter
openrouter-byok
$0
$0
Unavailable
Usage analytics

Loading usage…

API & code
Uptime & Health
No uptime data yet

These providers haven't been health-probed for this model yet. The router still routes around upstreams that fail live requests — uptime fills in once probe history accrues.

Share cards
Llama Guard 4 12B share card
Llama Guard 4 12B
DeepInfra upstream share card
DeepInfra upstream
OpenRouter upstream share card
OpenRouter upstream
Credits
Use your own key

Run Llama Guard 4 12B on your own key — your requests are billed by the provider. Pool callers pay AnyRouter credits.

No BYOK keys configured for this model yet.

Share a key with the pool to earn credits for every request it serves, covering your plan cost.

Text generation
Context length163,840 tokens
Max output16,384 tokens
ArchitectureTransformer
Categorymultimodal
ReleasedApr 29, 2025
Modalities
→
Capabilities