Plugins
4 pluginscurated
Run Agent Evaluation
Sets up evaluation framework, runs benchmarks, and produces comparative analysis of agent performance.
9 skills · plugin
@testdouble
Han Atlassian
Atlassian-facing extensions to the Han suite. Adds markdown-to-confluence, which publishes a local Markdown file to a user-specified Confluence page; project-documentation-to-confluence, which runs the han-documentation project-documentation skill and then publishes the result there; investigate-to-confluence, which runs the core investigate skill and publishes the resulting investigation report t
6 skills · plugin
@trailofbits
Trailmark
Builds multi-language source code graphs for security analysis: call graphs, attack surface mapping, blast radius, taint propagation, complexity hotspots, and entry point enumeration. Generates Mermaid diagrams (call graphs, class hierarchies, dependency maps, heatmaps). Compares code graph snapshots for structural diff and evolution analysis. Runs graph-informed mutation testing triage (genotoxic
10 skills · plugin
@testdouble
Han Communication
Foundational communication plugin for the Han suite. Owns the canonical readability standard, writing-voice profile, and explanation standard, the readability-guidance skill that surfaces the first two into a calling skill's context for in-voice drafting, the explanation-guidance skill that surfaces the third at the point a run talks to a person, the readability-editor agent that runs the adversar
3 skills · plugin
Results for “runs”
520 skillsJust Command Runner
Runs project-specific commands defined in a justfile, a Make-inspired syntax supporting parameters, dependencies, and multi-language recipes.
28
Full Test
Runs a complete testing pipeline covering full test suites, automated and manual tests, and logs execution metadata for the /evolve pipeline.
13
Generate
Regenerates a specific code layer (schemas, services, routes, hooks, ui) or all layers, then runs the corresponding validation checks.
1
Cron Jobs
Schedules and runs recurring commands or scripts using cron expressions, with support for intervals, retries, locking, and logging.
1
QA Cycle
Runs a QA validation cycle followed by bugfix until all criteria pass clean, with automated verification and re-validation.
4
Run
Run a single experiment iteration. Edit the target file, evaluate, keep or discard. Use when the user runs /ar:run or asks for one manual autoresearch iteration.
11
Outreach Sequencer
Send personalized cold email sequences via Gmail (self-managed loop), track replies, schedule follow-ups, and route booked meetings to Calendly. Runs daily on cron schedule.
0
Autoplan
Auto-review pipeline — reads the full CEO, design, eng, and DX review skills from disk and runs them sequentially with auto-decisions using 6 decision principles. (gstack)
0 · bundle
Marketing Ops
Routes marketing questions to the right specialist skill, orchestrates multi-skill campaigns, and runs cross-functional marketing audits.
20.4k · bundle
Ck Fix
Runs a structured bug-fix pipeline that scouts, diagnoses, fixes, reviews, and finalizes code changes with optional quality re-verification.
19 · bundle
Sr Brainstorm
Runs a structured five-round brainstorm to capture actors, features, scope, constraints, and business rules before writing a spec.
19 · bundle
Simplify Code
Runs three parallel reviewer agents over recent code changes, aggregates their findings, and applies the fixes worth applying.
2
Ck Fix
Runs a structured bug-fix pipeline: scouts evidence, diagnoses root cause, applies minimal fixes, and reviews before committing.
1 · bundle
Overnight Eval
Launches long-running evaluation batches in isolated tmux sessions with pre-flight verification, monitoring, and post-flight analysis for unattended runs.
0
Spike
Runs focused experiments to validate feasibility of an idea, saving artifacts to .planning/spikes/ and supporting both idea and frontier modes.
1 · bundle
Review
Runs independent AI CLI reviewers over phase plans and merges their feedback into a REVIEWS.md file.
1 · bundle
Static Analysis
Configures and runs static analysis and linting tools across multiple languages, integrating with CI/CD and security platforms.
4
Gws Gmail Forward
Forwards a Gmail message to new recipients, with options for body notes, attachments, CC/BCC, HTML formatting, drafts, and dry runs.
0
Cue
Routes research requests to appropriate modes and runs multi-agent deep research with evidence chains, plus optional monitoring.
1 · bundle
Github
Interact with GitHub issues, pull requests, Actions runs, and the GitHub API using the `gh` CLI.
5 · bundle
Interactive Command
Runs an interactive command in a separate window and waits for it to close. Invoked only when another skill explicitly calls for it, never on its own.
1 · bundle
Run
Run a single experiment iteration. Edit the target file, evaluate, keep or discard. Use when the user runs /ar:run or asks for one manual autoresearch iteration.
1
Pairing
Build work collaboratively in reviewable pieces, handing each piece back for review before starting the next, so the person stays in the lead and steers while the work happens instead of reviewing a finished result. Use when someone says to pair with them on something, asks to collaborate rather than direct, wants to review as it goes, or wants to guide the work piece by piece — on code, on a design decision, or on writing. For a test-first build it runs tdd, for restructuring it runs refactor, for an interface contract it runs design-an-api, and for plan work it runs iterative-plan-review or plan-implementation, each collaboratively; invoke any of those directly instead to run it straight through without pausing. Does not pace someone through code that already exists and builds nothing — use code-walkthrough. Does not explain, summarize, or research something instead of producing it — use code-overview or research.
218
Aiq Deploy
Installs, deploys, runs, validates, troubleshoots, and stops NVIDIA AI-Q Blueprint infrastructure for local or self-hosted servers.
2.2k · bundle
Trailmark Summary
Runs a Trailmark summary analysis on a codebase to auto-detect languages, count entry points, and list dependencies.
6k · bundle
Mysql Query Agent
Runs MySQL queries from an AI coding agent using the mysql2 Node.js library, with setup instructions for npm and TypeScript.
28
Apify Actor Runner
Runs Apify cloud actors for structured web scraping and exports datasets to S3, with input schema validation and webhook notifications.
28
Gemini
Runs Google's Gemini CLI in one-shot mode with a positional prompt, supporting model selection, JSON output, and extension management.
61
Aeo
Runs AEO audits on deployed sites, applies targeted fixes, validates JSON-LD schema, generates llms.txt files, and monitors changes over time.
32
Evaluate Edit
Runs regression evaluations comparing agent edits against human-approved golden projects, and registers new goldens after human approval.
3
Post Eval
Runs a post-batch analysis pipeline after an eval completes: verifies results, executes analysis scripts, refreshes dashboards, and writes a summary report.
0
Diagnose
Runs a disciplined debugging loop for hard bugs and performance regressions, covering reproduction, hypothesis testing, instrumentation, fixing, and regression testing.
0
Mlflow
Manages the machine learning lifecycle with experiment tracking, model versioning, reproducible runs, and deployment through the MLflow platform.
1
Bots
Builds and runs Telegram, Twitter/X, and WhatsApp bots for automated posting, engagement, and content distribution across platforms.
10
Job Scraper
Scrapes Danish job sites for new positions matching your profile. Deduplicates across runs. Triggers on: job scrape, find jobs, search jobs, new jobs, job search, scrape jobs, /scrape
0 · bundle
Arize Evaluator
Creates and runs LLM-as-judge evaluators on Arize, including managing tasks, column mappings, and continuous monitoring.
36.2k · bundle