Meta: Llama Guard 4 12B

by meta-llama

0 stars
Context 164K tokens
Modalities Text, Image → Text
Input Price $0.18 / million tokens
Output Price $0.18 / million tokens

Overview

Llama Guard 4 is a Llama 4 Scout-derived multimodal pretrained model, fine-tuned for content safety classification. Similar to previous versions, it can be used to classify content in both LLM inputs (prompt classification) and in LLM responses (response classification). It acts as an LLM—generating text in its output that indicates whether a given prompt or response is safe or unsafe, and if unsafe, it also lists the content categories violated. Llama Guard 4 was aligned to safeguard against the standardized MLCommons hazards taxonomy and designed to support multimodal Llama 4 capabilities. Specifically, it combines features from previous Llama Guard models, providing content moderation for English and multiple supported languages, along with enhanced capabilities to handle mixed text-and-image prompts, including multiple images. Additionally, Llama Guard 4 is integrated into the Llama Moderations API, extending robust safety classification to text and images.

Key Features

  • Multimodal capabilities (Text, Image → Text)
  • 164K tokens context window
  • API access available

Model Information

Developer:

meta-llama

Release Date:

April 30, 2025

Context Window:

164K tokens

Modalities:

Text, Image → Text

Pricing

Input Tokens $0.18 / million tokens
Output Tokens $0.18 / million tokens
Get API Key

Discussion

No comments yet. Be the first to share your thoughts about this model!