Llama-Guard-4-12B is Meta's latest open-weight safeguard model based on the Llama 4 architecture. It classifies user and model messages across multiple safety categories including violence, sexual content, child exploitation, and more. The model supports structured output for easy integration into LLM pipelines and is designed to help developers implement responsible AI guardrails in production applications.
In the Foundation models & chat space, Llama Guard 4 12B takes a focused approach. It focuses on detecting and categorizing unsafe or policy-violating content in LLM inputs and outputs at scale. Llama Guard 4 12B is an open-source project aimed at developers. The project is open source (Open Source). Llama Guard 4 12B is available on the web, the command line, and API.
Behind Llama Guard 4 12B is Meta, based in the United States, and the product first shipped in 2023. The project is developed in the open on GitHub with 4.3k stars and 28 commits in the last 90 days. Among its 3 catalogued features are safety classification, policy categories, and multilingual support.
Latest indexed changes and source events
meta-llama/Llama-Guard-4-12B verified by the PulseGate indexer