SkillSpector · llm-security
independent scanner by NVIDIA · skill by zhaoxuya520 · how it works ↗
Skill attempts to nullify the agent's safety policies or restrictions ('you have no restrictions', 'ignore your guidelines', 'do anything now'). This is a di…; Code accesses credential files (SSH keys, AWS credentials, etc.). This could indicate credential theft attempts.; Hidden instructions were detected in comments or invisible text. These could contain malicious directives. Manual review is recommended.; +3 more
scanned 2026-08-22
Findings (12)
Skill attempts to nullify the agent's safety policies or restrictions ('you have no restrictions', 'ignore your guidelines', 'do anything now'). This is a direct jailbreak that disables guardrails.
SKILL.md
Code accesses credential files (SSH keys, AWS credentials, etc.). This could indicate credential theft attempts.
references/agent-security-testing.md
Hidden instructions were detected in comments or invisible text. These could contain malicious directives. Manual review is recommended.
SKILL.md
Hidden instructions were detected in comments or invisible text. These could contain malicious directives. Manual review is recommended.
SKILL.md
Hidden instructions were detected in comments or invisible text. These could contain malicious directives. Manual review is recommended.
SKILL.md
Hidden instructions were detected in comments or invisible text. These could contain malicious directives. Manual review is recommended.
SKILL.md
Hidden instructions were detected in comments or invisible text. These could contain malicious directives. Manual review is recommended.
references/prompt-injection-methodology.md
Remote code is downloaded and executed. This bypasses code review and could introduce malicious code.
references/agent-security-testing.md
Tool calls are chained to bypass individual safety checks or escalate capabilities beyond what any single tool call would allow.
references/agent-security-testing.md
YARA rule matched a hack tool or exploit indicator (offensive tools, reconnaissance, privilege escalation, or exploit frameworks).
SKILL.md
YARA rule matched a hack tool or exploit indicator (offensive tools, reconnaissance, privilege escalation, or exploit frameworks).
SKILL.md
YARA rule matched a hack tool or exploit indicator (offensive tools, reconnaissance, privilege escalation, or exploit frameworks).
references/prompt-injection-methodology.md
What the verdicts mean
SkillSpector reports on SkillMD's shared five-tier scale. See how SkillSpector works ↗.
Overall severity LOW (risk score in the safe range)
Overall severity MEDIUM
Overall severity HIGH
Overall severity CRITICAL
Scan could not complete