Skip to main content
Toolsafety-alignmentv1.0

Llama Guard 3

by Meta · open-source · Last verified 2026-04-24

Meta's fine-tuned LLM classifier for detecting unsafe content in chat.

https://ai.meta.com/research/publications/llama-guard-llm-based-input-output-safeguard-for-human-ai-conversations/
D
DPoor
Adoption: C+Quality: B+Freshness: ACitations: FEngagement: F

Specifications

License
Open Source
Pricing
open-source
Capabilities
Integrations
Use Cases
API Available
No
SDK Languages
Tags
safety, moderation, meta, classifier
Added
2026-04-24
Completeness
73%

Index Score

34
Adoption
50
Quality
70
Freshness
80
Citations
0
Engagement
0

Need this tool deployed for your team?

Get a Custom Setup

Explore the full AI ecosystem on Agents as a Service