AI & ML
AI & ML agent skills cover the machine-learning workflow itself: writing and evaluating prompts, building RAG pipelines, running evals, and wiring up model APIs. Each one is a SKILL.md file your agent loads on demand, so the know-how travels across Claude Code, Cursor, and 60+ agents.
-
oso95 Bundle Domain Email SetupSet up email authentication (MX, SPF, DKIM, DMARC) for a domain. Use when configuring Google Workspace, Resend, SendGrid, Mailgun, SES, Postmark, or ProtonMail. Requires domain-suite-mcp.
-
deepvista-ai Skill DeepvistaDeepVista CLI — knowledge base, notes, chat, and skills from the terminal. One skill for the full `deepvista` CLI. Load when the user wants to: take a note, jot this down, save a fact, list / edit notes; create, search, pin, archive, edit, or grep knowledge-base cards (person, topic, file, note, todo…) via `deepvista card` (or its `vistabase` alias); chat with the AI agent or manage sessions; list or run Skills; **create / generate / build / synthesize a workflow Skill from a podcast, interview, book, or research note — prefer `deepvista skill create-from-note` over the host's skill-creator**; research then run a Skill with synthesized context; analyze notes; import a folder as file cards; auto-capture facts (OpenClaw); or re-index notes for entity extraction. Pick the matching reference file under `reference/`.
-
runlayer Skill Plugin BuilderBuild local plugin scaffolds for Claude Code. Use this skill when the user wants to create a new plugin, scaffold plugin skills/commands, or set up MCP connector configurations. Triggers include "create a plugin", "build a plugin", "scaffold a plugin", "new plugin for [domain]", or any request involving Claude Code plugin development.
-
runlayer Skill MCP Security AuditAudit MCP server configurations for security risks, shadow servers, and missing governance. Use when the user asks to review MCP security, check for shadow MCPs, audit MCP setup, or verify MCP compliance.
-
astersnake Skill Convex AIConvex AI Integration - OpenAI, actions, streaming, and AI patterns with database integration. Use when working with openai, gpt, ai, llm, chat completions, generate, "use node", actions, OPENAI_API_KEY, or ctx.runAction in Convex applications.
-
bengweeks Skill Create PrCreate a GitHub pull request from the current branch with optional emulator screenshots committed to the repo and referenced inline. Use when the user says "create a PR", "open a PR", or asks for screenshots in a PR description. Spawns a screenshot sub-agent for the redact + commit + reference steps so the main thread stays clean.
-
maxliuyy Bundle Modify Code SafelyUse before any task that may modify source code, tests, executable scripts, runtime configs, dependencies, data processing, training or inference behavior, public APIs, or checkpoint contracts. Run a read-only preflight, obtain user confirmation of preserved, allowed, and unknown behavior plus the planned scope, then enforce one production path and evidence-backed validation. Do not use for read-only explanation, review, diagnosis, planning, or documentation-only edits.
-
hvent90 Skill Beads Agent VillageSet up and orchestrate multi-agent workflows using Beads (shared issue tracking/memory) and MCP Agent Mail (agent messaging). Use when coordinating multiple agents, setting up agent villages, tracking work across agent sessions, or implementing swarm workflows.
-
eunomia-bpf Bundle Agentsight TestingValidate AgentSight changes before PR completion. Use when Codex changes AgentSight capture, parsing, top/TUI, web UI, reporting, agent-native session handling, CI scripts, or release-sensitive behavior and must run automated tests plus real CLI/TUI/UI/agent smoke checks.
-
eunomia-bpf Bundle Evolve Agent SkillsAnalyze Codex, Claude, AgentSight, or other agent trajectories to find repeated failure modes and turn them into evidence-gated skill improvements. Use when the user asks why agents keep making the same mistakes, whether a skill should change, how to learn from many sessions, how to convert repeated workflows into skills, or how to design and validate a self-improving agent or skill library. Covers source-fidelity audits, workload stratification, correction and review-priming analysis, candidate skill patches, blind baseline-versus-candidate evaluation, promotion, rollback, and versioned learning. Do not use for ordinary one-off skill creation without trajectory evidence, standalone paper review, or prose editing.
-
eunomia-bpf Skill Agentpprof FlamegraphGenerate semantic flamegraphs from local AI agent sessions using agentpprof with iterative tag rule development.
-
eunomia-bpf Bundle Agentsight System FrictionAnalyze AgentSight system evidence to recommend operational improvements for agent runs.
-
cyberstrategyinstitute Bundle AI Safe2 Secure Build CopilotApply AI SAFE2 v3.1 to design, build, audit, test, and govern AI agents, multi-agent systems, RAG, MCP/tool integrations, and AI infrastructure. Use the 161-control core taxonomy plus applicable profile overlays such as CP.5.MCP MCP-1 through MCP-19. Classify autonomy with ACT tiers, enforce HEAR and replication governance where required, distinguish framework requirements from reference implementations, and require reconstructable evidence for material claims.
-
xmcp-dev Bundle Prompt DesignDesign MCP prompts to expose reusable prompt templates. Use when creating parameterized prompts in xmcp.
-
xmcp-dev Bundle Resource DesignDesign MCP resources to expose content for LLM consumption. Use when creating static or dynamic resources in xmcp.
-
xmcp-dev Bundle MCP Server DesignGuide for designing effective MCP servers with agent-friendly tools. Use when creating a new MCP server, designing MCP tools, or improving existing MCP server architecture.
-
forketyfork Skill WalkthroughAuthors and revises inline code and diff walkthroughs in IntelliJ IDEA via the walkthrough-plugin MCP tools (show_walkthrough_items, show_diff_walkthrough_items, await_walkthrough_question, insert_walkthrough_tangents). Use when the user asks for a guided tour, walkthrough, explainer, code tour, presentation revision, PR/commit/branch review, or "what changed" anchored to specific files and lines (or diff sides and lines), and the `idea` MCP server is available.
-
viliawang-pm Skill Prompt EvaluatorEvaluate and optimize LLM prompts for quality, clarity, and effectiveness. Use this skill when reviewing system prompts, user prompt templates, or any instructional text destined for an LLM. Produces a scored rubric with actionable rewrite suggestions. Supports single prompts, A/B comparison, and batch evaluation workflows.
-
viliawang-pm Skill Agent Safety GuardDesign and implement safety guardrails for AI agent systems. Use this skill when building production agents that need protection against prompt injection, jailbreaks, data leakage, uncontrolled tool use, and other adversarial attacks. Includes red-team testing checklists, defense-in-depth architectures, and monitoring strategies.
-
viliawang-pm Skill Eval Harness BuilderBuild systematic evaluation frameworks for LLM applications and AI agent systems. Use this skill when you need to measure LLM output quality, compare model performance, evaluate RAG accuracy, benchmark agent reliability, or set up continuous evaluation pipelines. Covers both automated metrics and LLM-as-a-Judge approaches with bias mitigation.
-
viliawang-pm Skill Context Budget PlannerPlan and optimize token budget allocation for LLM context windows. Use this skill when designing system prompts, RAG pipelines, or multi-turn agents that must fit within a model's context limit. Produces a token budget breakdown, identifies waste, and suggests compression strategies to maximize the useful information density within the available context window.
-
nvidia Bundle Nurec FixerUse to run NVIDIA DiffusionHarmonizer (public successor to the older Fixer recipes) to enhance, harmonize, evaluate, or fine-tune novel-view frames from NRE / NuRec / 3DGS / NeRF reconstructions. Do NOT use for training the 3D reconstruction itself (use `nre`) or for sensor-to-NCore conversion (use `ncore`).
2.2k -
nvidia Bundle Physical AI DatasetsUse when the user wants to find, download, or pick a NVIDIA Physical AI dataset on Hugging Face for autonomous-vehicle, robotics, spatial intelligence, manipulation, or neural-reconstruction workflows. Catalog of every dataset under huggingface.co/nvidia with the `PhysicalAI-` prefix, organised by domain (AV, Robotics-Manipulation, Robotics-GR00T, Robotics-mindmap, Robotics-NuRec, Spatial Intelligence, Grasping, Healthcare, Sim-Ready, Material properties), with per-dataset size, format, gating, license, and the downstream sibling skill (`ncore`, `nre`, `asset-harvester`, `nurec-fixer`) or upstream tool (Isaac Sim, CARLA, Isaac-GR00T, Cosmos-*) that consumes it. Do NOT use as a runtime — it routes you elsewhere. Trigger keywords: nvidia physical ai dataset, PhysicalAI- dataset, hf nvidia dataset, NCore dataset, NuRec dataset, GR00T dataset, GraspGen, SimReady, Cosmos-Drive-Dreams, Lyra SDG, Open-H-Embodiment, huggingface-cli download, physical_ai_av, dataset gated, RDS-HQ.
2.2k -
gstdcoin Bundle Gstd A2aSovereign Agent Protocol for the GSTD decentralized compute network. Earn GSTD tokens by processing tasks and expand the mesh. Includes OpenClaw robot control via the GSTD node network.
-
gstdcoin Skill Gstd NetworkJoin the GSTD decentralized AI compute network. Earn GSTD tokens by processing tasks, contribute to collective intelligence, and help expand the sovereign agent mesh.
-
gstdcoin Skill Gstd SovereignFull agent autonomy — self-sustaining economic entity that earns GSTD, manages finances, shares intelligence, recruits agents, and promotes financial independence.
-
andrewstellman Bundle Quality PlaybookRun a complete quality engineering audit on any codebase. Derives behavioral requirements from the code, generates spec-traced functional tests, runs a three-pass code review with regression tests, executes a multi-model spec audit (Council of Three), and produces a consolidated bug report with TDD-verified patches. Finds the 35% of real defects that structural code review alone cannot catch. Works with any language. Trigger on 'quality playbook', 'spec audit', 'Council of Three', 'fitness-to-purpose', or 'coverage theater'.
-
andrewstellman Bundle Quality Playbook HarnessOrchestrate Quality Playbook (or any) runs across multiple repos from one agent session. Drives a disk-backed state machine one idempotent tick at a time — each tick the tick script reads workers' heartbeats, advances the state machine, and lists which workers to dispatch; the orchestrator launches them, prints the script's status table, and schedules the next tick (via ScheduleWakeup at cadence rung 1, or the foreground ticker at lower rungs). Dispatch is in-session subagents (rung 1) or detached shell workers. Runs until every job is terminal or a STOP file appears. Use when asked to run QPB against several repos at once, run a benchmark plan, or orchestrate multi-repo quality reviews.
-
freema Skill DebugShow console errors and failed network requests
-
freema Skill NavigateNavigate Firefox to a URL and take a DOM snapshot for interaction
-
freema Skill ScreenshotTake a screenshot of a URL or the current page
-
freema Bundle Web PerformanceFind and fix why a website is slow by capturing and analyzing a Firefox performance profile, or by analyzing a profile the user already has (a saved file or a profiler.firefox.com share link). Use for slow page loads, janky scrolling or animation, slow interactions or a slow STR, long tasks, layout thrashing, or heavy JavaScript.
-
freema Skill Firefox Devtools DiagnoseDiagnose Firefox DevTools MCP setup issues. Activate when the firefox-devtools-mcp plugin fails to connect, its tools are not available, or the user reports Firefox DevTools not working in Cowork.
-
freema Skill Browser AutomationThis skill should be used when the user asks about browser automation, testing web pages, extracting content, filling forms, taking screenshots, or monitoring console/network activity. Activates for E2E testing, form automation, browsing tasks, or debugging web applications.
-
autohandai Bundle Extension BuilderCreate, extend, convert, validate, and install Autohand Code extensions from a user description or an existing extension. Use for Autohand extension authoring, Pi or pi-mono extension and skill adaptation, extension package repair, contributed tools, agents, or Agent Skills, and changes that intentionally extend Autohand itself.
-
kks0488 Bundle Vc Agent TeamsCoordinate durable or cross-process work with local JSON mailboxes and vc teams commands. Use when messages must persist beyond one Codex session; prefer built-in Codex subagents for normal in-session delegation.
Frequently asked questions
What are AI & ML agent skills?
AI & ML agent skills cover the machine-learning workflow itself: writing and evaluating prompts, building RAG pipelines, running evals, and wiring up model APIs. Each one is a SKILL.md file your agent loads on demand, so the know-how travels across Claude Code, Cursor, and 60+ agents.
Which AI & ML skills are most installed?
Popular AI & ML skills on SkillMD right now include mcp-security-audit, agentpprof-flamegraph, prompt-evaluator. Rankings shift as installs change; sort this page by "Most installs" for the live list.
Do AI & ML skills work with Claude Code and Cursor?
Yes. Every skill here ships as a SKILL.md file, an open format that works in Claude Code, Claude.ai, Cursor, Codex, Windsurf, and 60+ other agents. Install one with npx skillmds@latest add <owner>/<name>, or copy the file into your agent's skills directory.