SkillSpector · prompt-guard
independent scanner by NVIDIA · skill by Orchestra Research · how it works ↗
Skill attempts to nullify the agent's safety policies or restrictions ('you have no restrictions', 'ignore your guidelines', 'do anything now'). This is a direc; This pattern attempts to override system instructions or ignore safety constraints. Without LLM analysis, manual review is recommended.; YARA rule matched a hack tool or exploit indicator (offensive tools, reconnaissance, privilege escalation, or exploit frameworks).
scanned 2026-07-07
Findings (3)
Skill attempts to nullify the agent's safety policies or restrictions ('you have no restrictions', 'ignore your guidelines', 'do anything now'). This is a direc
SKILL.md
This pattern attempts to override system instructions or ignore safety constraints. Without LLM analysis, manual review is recommended.
SKILL.md
YARA rule matched a hack tool or exploit indicator (offensive tools, reconnaissance, privilege escalation, or exploit frameworks).
SKILL.md
What the verdicts mean
SkillSpector reports on SkillMD's shared five-tier scale. See how SkillSpector works ↗.
Overall severity LOW (risk score in the safe range)
Overall severity MEDIUM
Overall severity HIGH
Overall severity CRITICAL
Scan could not complete