red-teaming-llms-with-garak

mukul975/red-teaming-llms-with-garak · Agent Skill (multi-file)

by mukul975 · bundle

Published · Last updated


Run NVIDIA garak probe suites against an LLM endpoint to test for jailbreaks, prompt injection, data leakage, and toxic generation, then interpret the hit-rate report for triage and reporting.

SKILL.md

Files

This skill is a package of 5 files. Install with the command above, or download the folder.

  • 📄SKILL.md entry
  • 📁references
  • 📄api-reference.md 3.0 KB
  • 📄standards.md 1.4 KB
  • 📁scripts
  • ⚙️agent.py 5.3 KB
  • 📄LICENSE 11.0 KB

Related

  1. testing-for-system-prompt-leakage · mukul975 bundle
    Test LLM applications for system prompt leakage using manual payloads, garak, and Promptfoo to extract embedded secrets and routing logic.
    24.6k
    repo stars
  2. ai-security · alirezarezvani bundle
    Assess AI/ML systems for prompt injection, jailbreak vulnerabilities, model inversion risk, data poisoning exposure, and agent tool abuse, with MITRE ATLAS mapping and guardrail recommendations.
    20.4k
    repo stars
  3. orchestrating-llm-attacks-with-pyrit · mukul975 bundle
    Automate multi-turn adversarial conversations against LLM agents using Microsoft PyRIT, including Crescendo and Tree-of-Attacks-with-Pruning (TAP) attack chains with scorer feedback loops.
    24.6k
    repo stars
  4. llm-security · zhaoxuya520 bundle
    Conduct authorized security assessments of LLM applications and AI agents, covering prompt injection, tool abuse, RAG exposure, memory poisoning, and model supply-chain risks.
    12.8k
    repo stars
  5. godmode · lord1egypt
    Bypasses safety filters on API-served LLMs using jailbreak templates, input obfuscation, and multi-model racing.
    2
    repo stars
  6. testing-prompt-injection-in-rag-pipelines · mukul975 bundle
    Probe RAG applications for prompt injection via poisoned retrieved context and embedding manipulation.
    24.6k
    repo stars

Frequently asked questions

How do I install the red-teaming-llms-with-garak skill?

Run npx skillmds add mukul975/red-teaming-llms-with-garak in your terminal (requires Node.js), paste this page's agent-chat prompt into Claude, Cursor, or any MCP-connected agent, or download the SKILL.md file and copy it into your agent's skills directory.

What does the red-teaming-llms-with-garak skill do?

Run NVIDIA garak probe suites against an LLM endpoint to test for jailbreaks, prompt injection, data leakage, and toxic generation, then interpret the hit-rate report for triage and reporting. It is listed under AI & ML, Security, Penetration Testing, Vulnerability Scanning on SkillMD.

Is red-teaming-llms-with-garak safe to use?

SkillMD's automated safety review verdict for this skill is CAUTION. Independent scanners report: SkillSpector: WARNING, Skill Scanner: PASS. Capability flags: executes scripts, makes network calls, reads secrets. SkillMD never runs a skill's scripts for you; review the SKILL.md before installing.

Which AI agents work with red-teaming-llms-with-garak?

This skill is tagged as working with Claude Code, Claude.ai, OpenAI Codex. SKILL.md is an open format, so most agents that read a skills directory can load it too.

Is red-teaming-llms-with-garak free to use?

Yes. Installing skills from SkillMD is free. This skill is licensed under Apache-2.

Who published red-teaming-llms-with-garak?

mukul975 (@mukul975) published this skill. Their other Agent Skills are listed on their SkillMD profile.