Eval Dataset Adversarial Prompts

Use when running or extending the adversarial prompt benchmark dataset that tests the legal AI system's robustness against jailbreaks, out-of-scope requests, unauthorized-practice attempts, privacy violations, and hallucination bait. This dataset catches the most expensive failure modes and must be run on every model deployment.

sboghossian Updated

File contents

sboghossian/mini-claude-for-legal/tree/main/skills/eval/eval-dataset-adversarial-prompts commit 513bc9d3b2

Frequently asked questions

npx skillmds@latest add sboghossian-mini-claude-for-legal/eval-dataset-adversarial-prompts