TLDRocket
Sign in

Welcoming Llama Guard 4 on Hugging Face Hub

Hugging Face Blog

Meta released Llama Guard 4, a 12-billion parameter multimodal safety model designed to detect unsafe content in both images and text across input prompts and model-generated outputs. The model can run on a single GPU with 24GB of VRAM and classifies 14 hazard types from the MLCommons taxonomy, improving recall by 4 percentage points and F1-score by 8 points compared to its predecessor Llama Guard 3. The release enables flexible content moderation pipelines where user inputs are filtered before reaching language models and generated responses can be reviewed for safety before delivery.

Related stories

The daily briefing

Every AI story that matters, in your inbox by 8am.

TLDRocket reads 60+ sources, removes duplicate coverage, and summarises the day in two minutes. Free, no spam, unsubscribe anytime.