← back to defending-llms-with-guardrails

SkillSpector · defending-llms-with-guardrails

independent scanner by NVIDIA · skill by mukul975 · how it works ↗

WARNINGmax severity: HIGHrisk score: 71

This pattern attempts to override system instructions or ignore safety constraints. Without LLM analysis, manual review is recommended.; Skill contains instructions that could directly expose system prompts, internal rules, or hidden instructions to users or external parties.; YARA rule matched a hack tool or exploit indicator (offensive tools, reconnaissance, privilege escalation, or exploit frameworks).

scanned 2026-07-07

Findings (4)

HIGHPrompt Injectionconfidence: 0.8

This pattern attempts to override system instructions or ignore safety constraints. Without LLM analysis, manual review is recommended.

SKILL.md

HIGHSystem Prompt Leakageconfidence: 0.85

Skill contains instructions that could directly expose system prompts, internal rules, or hidden instructions to users or external parties.

SKILL.md

HIGHSystem Prompt Leakageconfidence: 0.85

Skill contains instructions that could directly expose system prompts, internal rules, or hidden instructions to users or external parties.

SKILL.md

HIGHYARA Matchconfidence: 0.8

YARA rule matched a hack tool or exploit indicator (offensive tools, reconnaissance, privilege escalation, or exploit frameworks).

SKILL.md

What the verdicts mean

SkillSpector reports on SkillMD's shared five-tier scale. See how SkillSpector works ↗.

PASS

Overall severity LOW (risk score in the safe range)

CAUTION

Overall severity MEDIUM

WARNINGthis skill

Overall severity HIGH

FAIL

Overall severity CRITICAL

INCONCLUSIVE

Scan could not complete