Purple Llama
Meta's collection of trust and safety tools for language models, including the CyberSecEval benchmark suite and the Llama Guard family of input and output classifiers.
Services
- Safety benchmarks
- Input/output classifiers
- Prompt injection tests
- Cybersecurity evals
Highlights
Source: Official GitHub repository. Last checked 2026-09-18. Spot an error? Tell us.