Llama Guard 4 12B

Open multimodal safety classifier for prompts and responses.

Llama Guard 4 is derived from Llama 4 Scout and fine-tuned for content-safety classification. It classifies both prompts (input) and responses (output) — including images — against a hazard taxonomy, returning safe/unsafe plus the violated categories.

Provider
Meta
Type
Safety classifier
Released
Apr 29, 2025
Context window
164K tokens
Max output
16K tokens
Price
$0.18 input / $0.18 output per 1M tokens
Input
text, image
Output
text
Open weights
Yes
License
Llama 4 Community License

Best for

Strengths

More from Meta

Tools that use it

Sources