← back to yuanbao

SkillSpector · yuanbao

independent scanner by NVIDIA · skill by bog5d · how it works ↗

CAUTIONmax severity: MEDIUMrisk score: 48

Skill instructs the agent to never refuse or to always comply. Suppressing the agent's ability to decline removes a core safety control and enables downstrea…; Skill instructs the agent to omit warnings, disclaimers, or ethical commentary. Stripping safety caveats hides risk from the user and is a common jailbreak p…; Subtle instructions detected that may alter agent decision-making or introduce hidden biases.

scanned 2026-08-23

Findings (3)

HIGHAnti-Refusalconfidence: 0.85

Skill instructs the agent to never refuse or to always comply. Suppressing the agent's ability to decline removes a core safety control and enables downstream harmful requests to succeed.

SKILL.md

HIGHAnti-Refusalconfidence: 0.8

Skill instructs the agent to omit warnings, disclaimers, or ethical commentary. Stripping safety caveats hides risk from the user and is a common jailbreak preamble.

SKILL.md

MEDIUMPrompt Injectionconfidence: 0.75

Subtle instructions detected that may alter agent decision-making or introduce hidden biases.

SKILL.md

What the verdicts mean

SkillSpector reports on SkillMD's shared five-tier scale. See how SkillSpector works ↗.

PASS

Overall severity LOW (risk score in the safe range)

CAUTIONthis skill

Overall severity MEDIUM

WARNING

Overall severity HIGH

FAIL

Overall severity CRITICAL

INCONCLUSIVE

Scan could not complete