Llamaguard

Meta's 7-8B specialized moderation model for LLM input/output filtering. 6 safety categories - violence/hate, sexual content, weapons, substances, self-harm, criminal planning. 94-95% accuracy. Deploy with vLLM, HuggingFace, Sagemaker. Integrates with NeMo Guardrails.

synthetic-sciences 40be590 8.3 KB Updated

File contents

synthetic-sciences/openscience/tree/main/backend/cli/skills/llm-tools/llamaguard commit 40be59060c

Frequently asked questions

npx skillmds@latest add synthetic-sciences/llamaguard