AI Security

by Alireza Rezvani alirezarezvani/ai-security multi-file Updated


Assess AI/ML systems for prompt injection, jailbreak vulnerabilities, model inversion risk, data poisoning exposure, and agent tool abuse, with MITRE ATLAS mapping and guardrail recommendations.

SKILL.md

Related

  1. Testing For System Prompt Leakage · mukul975 bundle
    Test LLM applications for system prompt leakage using manual payloads, garak, and Promptfoo to extract embedded secrets and routing logic.
    24.6k
    repo stars
  2. Orchestrating LLM Attacks With Pyrit · mukul975 bundle
    Automate multi-turn adversarial conversations against LLM agents using Microsoft PyRIT, including Crescendo and Tree-of-Attacks-with-Pruning (TAP) attack chains with scorer feedback loops.
    24.6k
    repo stars
  3. Detecting Indirect Prompt Injection · mukul975 bundle
    Detect and defend against prompt injection hidden in documents, web pages, and images consumed by an agent.
    24.6k
    repo stars
  4. LLM Security · zhaoxuya520 bundle
    Conduct authorized security assessments of LLM applications and AI agents, covering prompt injection, tool abuse, RAG exposure, memory poisoning, and model supply-chain risks.
    12.8k
    repo stars
  5. Red Teaming Llms With Garak · mukul975 bundle
    Run NVIDIA garak probe suites against an LLM endpoint to test for jailbreaks, prompt injection, data leakage, and toxic generation, then interpret the hit-rate report for triage and reporting.
    24.6k
    repo stars
  6. Defending Llms With Guardrails · mukul975 bundle
    Deploy Llama Guard, NeMo Guardrails, and LLM Guard as runtime input/output scanners to block jailbreaks, prompt injection, and toxic content in production LLM applications.
    24.6k
    repo stars

Frequently asked questions

How do I install the AI Security skill?

Run npx skillmds add alirezarezvani/ai-security in your terminal (requires Node.js), paste this page's agent-chat prompt into Claude, Cursor, or any MCP-connected agent, or download the SKILL.md file and copy it into your agent's skills directory.

What does the AI Security skill do?

Assess AI/ML systems for prompt injection, jailbreak vulnerabilities, model inversion risk, data poisoning exposure, and agent tool abuse, with MITRE ATLAS mapping and guardrail recommendations. It is listed under Security, AI & ML, Penetration Testing, Prompt Engineering, Vulnerability Scanning on SkillMD.

Is AI Security safe to use?

SkillMD's automated safety review verdict for this skill is PASS. Independent scanners report: SkillSpector: FAIL, Skill Scanner: WARNING. Capability flags: executes scripts, makes network calls, reads secrets. SkillMD never runs a skill's scripts for you; review the SKILL.md before installing.

Which AI agents work with AI Security?

This skill is tagged as working with Claude Code, Claude.ai, OpenAI Codex. SKILL.md is an open format, so most agents that read a skills directory can load it too.

Is AI Security free to use?

Yes. Installing skills from SkillMD is free, and the skill stays under its author's original license.

Who published AI Security?

Alireza Rezvani (@alirezarezvani) published this skill. Their other Agent Skills are listed on their SkillMD profile.