SkillSpector · prompt-guard
independent scanner by NVIDIA · skill by qcmuu · how it works ↗
Skill attempts to nullify the agent's safety policies or restrictions ('you have no restrictions', 'ignore your guidelines', 'do anything now'). This is a di…; This pattern attempts to override system instructions or ignore safety constraints. Without LLM analysis, manual review is recommended.; YARA rule matched a hack tool or exploit indicator (offensive tools, reconnaissance, privilege escalation, or exploit frameworks).
scanned 2026-08-23
Findings (3)
Skill attempts to nullify the agent's safety policies or restrictions ('you have no restrictions', 'ignore your guidelines', 'do anything now'). This is a direct jailbreak that disables guardrails.
SKILL.md
This pattern attempts to override system instructions or ignore safety constraints. Without LLM analysis, manual review is recommended.
SKILL.md
YARA rule matched a hack tool or exploit indicator (offensive tools, reconnaissance, privilege escalation, or exploit frameworks).
SKILL.md
What the verdicts mean
SkillSpector reports on SkillMD's shared five-tier scale. See how SkillSpector works ↗.
Overall severity LOW (risk score in the safe range)
Overall severity MEDIUM
Overall severity HIGH
Overall severity CRITICAL
Scan could not complete