← back to academic-pipeline

SkillSpector · academic-pipeline

independent scanner by NVIDIA · skill by richardnguyen0715 · how it works ↗

FAILmax severity: CRITICALrisk score: 100

Skill instructs the agent to omit warnings, disclaimers, or ethical commentary. Stripping safety caveats hides risk from the user and is a common jailbreak p…; Skill enables autonomous high-impact decisions without human-in-the-loop verification. Critical operations (destructive commands, financial transactions, dat…; Skill manipulates agent memory, state, or stored context. Memory corruption can alter personality, override safety rules, or cause unpredictable behavior.; +2 more

scanned 2026-08-23

Findings (10)

HIGHAnti-Refusalconfidence: 0.85

Skill instructs the agent to omit warnings, disclaimers, or ethical commentary. Stripping safety caveats hides risk from the user and is a common jailbreak preamble.

agents/collaboration_depth_agent.md

MEDIUMExcessive Agencyconfidence: 0.8

Skill enables autonomous high-impact decisions without human-in-the-loop verification. Critical operations (destructive commands, financial transactions, data deletion) should require explicit user confirmation.

agents/pipeline_orchestrator_agent.md

HIGHMemory Poisoningconfidence: 0.8

Skill manipulates agent memory, state, or stored context. Memory corruption can alter personality, override safety rules, or cause unpredictable behavior.

agents/pipeline_orchestrator_agent.md

HIGHMemory Poisoningconfidence: 0.24

Skill manipulates agent memory, state, or stored context. Memory corruption can alter personality, override safety rules, or cause unpredictable behavior.

references/passport_as_reset_boundary.md

HIGHPrompt Injectionconfidence: 0.7

Hidden instructions were detected in comments or invisible text. These could contain malicious directives. Manual review is recommended.

agents/claim_ref_alignment_audit_agent.md

HIGHPrompt Injectionconfidence: 0.7

Hidden instructions were detected in comments or invisible text. These could contain malicious directives. Manual review is recommended.

agents/claim_ref_alignment_audit_agent.md

HIGHPrompt Injectionconfidence: 0.7

Hidden instructions were detected in comments or invisible text. These could contain malicious directives. Manual review is recommended.

agents/claim_ref_alignment_audit_agent.md

HIGHPrompt Injectionconfidence: 0.7

Hidden instructions were detected in comments or invisible text. These could contain malicious directives. Manual review is recommended.

agents/pipeline_orchestrator_agent.md

HIGHPrompt Injectionconfidence: 0.7

Hidden instructions were detected in comments or invisible text. These could contain malicious directives. Manual review is recommended.

agents/pipeline_orchestrator_agent.md

HIGHSystem Prompt Leakageconfidence: 0.85

Skill contains instructions that could directly expose system prompts, internal rules, or hidden instructions to users or external parties.

agents/state_tracker_agent.md

What the verdicts mean

SkillSpector reports on SkillMD's shared five-tier scale. See how SkillSpector works ↗.

PASS

Overall severity LOW (risk score in the safe range)

CAUTION

Overall severity MEDIUM

WARNING

Overall severity HIGH

FAILthis skill

Overall severity CRITICAL

INCONCLUSIVE

Scan could not complete