Default output: return only the result, blockers, and required evidence. Omit preambles, process narration, repeated context, confidence scores, and follow-up offers. Use at most five bullets unless a required artifact or schema needs more.
Using Agent Skills
Overview
This is the meta-skill for discovering and invoking the right skill, command, or agent for any task. It includes a decision tree for task-to-skill mapping and core operating behaviors that all agents should follow.
When to Use
- Starting a new session and unsure which skill applies
- Task doesn't clearly match a single skill
- Need to understand the full capability landscape
- Debugging why an agent isn't following expected workflows
Decision Tree: Task to Skill Mapping
What are you trying to do?
DEFINE (understand the problem)
→ Writing requirements before coding?
→ spec-driven-development
PLAN (break down the work)
→ Decomposing a spec into implementable tasks?
→ planning-and-task-breakdown
→ Researching before planning?
→ comprehensive-research
→ Stress-testing a plan or design before committing?
→ grill-me
→ Grilling a plan and persisting the decisions as a document?
→ grill-with-docs
LEARN (teach and grow)
→ Interactive lessons over multiple sessions with a file-backed workspace?
→ teach
→ Week-by-week roadmap and milestones without lesson artifacts?
→ create-learning-path
→ Team knowledge base or domain map (not personal tutoring)?
→ knowledge-architecture
BUILD (write the code)
→ Implementing a feature across multiple files?
→ incremental-implementation
→ Writing tests first?
→ test-driven-development
→ Need the right documentation?
→ source-driven-development
VERIFY (prove it works)
→ Testing in the browser?
→ browser-testing-with-devtools
→ Something is broken?
→ debugging-and-error-recovery
REVIEW (improve code health)
→ Reviewing code before merge?
→ code-review-and-quality
→ Full repo health check or improvement plan?
→ repo-audit (or /ai-eng/repo-audit)
→ Code works but is hard to read?
→ code-simplification
→ Security concerns?
→ security-and-hardening
SHIP (deploy safely)
→ Preparing for production?
→ shipping-and-launch
→ Setting up CI/CD?
→ ci-cd-and-automation
SPECIALIZED
→ Building knowledge graphs?
→ graph-rag
→ Creating plugins/commands/skills?
→ plugin-dev
→ Building a harness workflow runner (Cursor, Claude, OpenCode, OpenAI, Pi)?
→ agents-sdk-dev (unified — specify harness)
→ See agents/research-runner/ for reference implementations
→ Gemini workflow adapter (not shipped yet)?
→ gemini-agent-sdk
→ Large cloud multi-agent tree (/orchestrate)?
→ orchestrate (planned; use agents-sdk-dev with harness=cursor until scripts/cli.ts exists)
→ Optimizing prompts?
→ prompt-refinement
→ Cleaning up AI verbosity?
→ text-cleanup
→ Managing git worktrees?
→ git-worktree
→ Deploying to Coolify?
→ coolify-deploy
Core Operating Behaviors
1. Surface Assumptions
Before acting, state what you're assuming:
- "I'm assuming this is a new feature, not a bug fix"
- "I'm assuming we're using the existing auth pattern"
- "I'm assuming the database schema doesn't need changes"
2. Manage Confusion
When uncertain, say so explicitly:
- "I'm not sure which pattern applies here"
- "This could be interpreted two ways"
- "I need more context about the expected behavior"
3. Push Back
Challenge requirements when needed:
- "This approach will create tight coupling. Consider..."
- "The spec says X, but the existing code does Y. Which is correct?"
- "This change is too large for one commit. Should we split it?"
4. Enforce Simplicity
Prefer simple solutions:
- "Can we solve this with a function instead of a class?"
- "Do we need a new dependency, or can we use what we have?"
- "Is there a simpler way to handle this edge case?"
5. Scope Discipline
Stay within boundaries:
- "This is out of scope for the current task"
- "I'll note this as a follow-up, not implement it now"
- "The spec covers A and B, but not C. Should I expand scope?"
6. Verify
Always prove correctness:
- "Seems right" is never sufficient
- Run tests, check build output, verify runtime behavior
- Provide evidence: passing tests, screenshots, logs
Anti-Rationalization Table
| Excuse |
Counter |
| "I'll add tests later" |
Tests are proof. Without them, you have no evidence it works. Write tests first or alongside. |
| "This is simple enough to skip the spec" |
Simple now, complex later. A 5-minute spec prevents 5-hour rewrites. |
| "I know this pattern, no need to check docs" |
Frameworks change. Source-driven development prevents subtle bugs from outdated knowledge. |
| "The code works, no need to simplify" |
Working code that's hard to read is technical debt. Future you (or teammates) will pay the cost. |
| "I'll fix the security issue in the next PR" |
Security issues are stop-the-line. Fix now or create a tracked issue with severity. |
| "This only affects one file, no review needed" |
Every change deserves review. Even small changes can have cascading effects. |
| "I'll document it later" |
Undocumented code is a time bomb. Document the why, not the what, while it's fresh. |
Composition Rules
- Agents do NOT invoke other agents — enforced by platform constraint
- Agents MAY invoke skills — skills are instructions, not agents
- The user or a slash command is the orchestrator — no router agents
- Skills auto-activate based on context — no manual skill selection needed
- Teams cannot nest — flat hierarchy only
References
references/orchestration-patterns.md — Endorsed and anti-pattern orchestration approaches
references/testing-patterns.md — Test structure, naming, mocking, examples
references/security-checklist.md — Pre-commit checks, OWASP Top 10, secrets management
references/performance-checklist.md — Core Web Vitals, frontend/backend checklists
references/accessibility-checklist.md — Keyboard nav, screen readers, WCAG 2.1 AA
1---2name: using-agent-skills3description: Discovery and invocation of all skills, commands, and agents. Decision tree for task-to-skill mapping. Core operating behaviors for agents. Use when unsure which skill or agent to use.4---56Default output: return only the result, blockers, and required evidence. Omit preambles, process narration, repeated context, confidence scores, and follow-up offers. Use at most five bullets unless a required artifact or schema needs more.78# Using Agent Skills910## Overview1112This is the meta-skill for discovering and invoking the right skill, command, or agent for any task. It includes a decision tree for task-to-skill mapping and core operating behaviors that all agents should follow.1314## When to Use1516- Starting a new session and unsure which skill applies17- Task doesn't clearly match a single skill18- Need to understand the full capability landscape19- Debugging why an agent isn't following expected workflows2021## Decision Tree: Task to Skill Mapping2223```24What are you trying to do?2526DEFINE (understand the problem)27 → Writing requirements before coding?28 → spec-driven-development2930PLAN (break down the work)31 → Decomposing a spec into implementable tasks?32 → planning-and-task-breakdown33 → Researching before planning?34 → comprehensive-research35 → Stress-testing a plan or design before committing?36 → grill-me37 → Grilling a plan and persisting the decisions as a document?38 → grill-with-docs3940LEARN (teach and grow)41 → Interactive lessons over multiple sessions with a file-backed workspace?42 → teach43 → Week-by-week roadmap and milestones without lesson artifacts?44 → create-learning-path45 → Team knowledge base or domain map (not personal tutoring)?46 → knowledge-architecture4748BUILD (write the code)49 → Implementing a feature across multiple files?50 → incremental-implementation51 → Writing tests first?52 → test-driven-development53 → Need the right documentation?54 → source-driven-development5556VERIFY (prove it works)57 → Testing in the browser?58 → browser-testing-with-devtools59 → Something is broken?60 → debugging-and-error-recovery6162REVIEW (improve code health)63 → Reviewing code before merge?64 → code-review-and-quality65 → Full repo health check or improvement plan?66 → repo-audit (or /ai-eng/repo-audit)67 → Code works but is hard to read?68 → code-simplification69 → Security concerns?70 → security-and-hardening7172SHIP (deploy safely)73 → Preparing for production?74 → shipping-and-launch75 → Setting up CI/CD?76 → ci-cd-and-automation7778SPECIALIZED79 → Building knowledge graphs?80 → graph-rag81 → Creating plugins/commands/skills?82 → plugin-dev83 → Building a harness workflow runner (Cursor, Claude, OpenCode, OpenAI, Pi)?84 → agents-sdk-dev (unified — specify harness)85 → See agents/research-runner/ for reference implementations86 → Gemini workflow adapter (not shipped yet)?87 → gemini-agent-sdk88 → Large cloud multi-agent tree (/orchestrate)?89 → orchestrate (planned; use agents-sdk-dev with harness=cursor until scripts/cli.ts exists)90 → Optimizing prompts?91 → prompt-refinement92 → Cleaning up AI verbosity?93 → text-cleanup94 → Managing git worktrees?95 → git-worktree96 → Deploying to Coolify?97 → coolify-deploy98```99100## Core Operating Behaviors101102### 1. Surface Assumptions103Before acting, state what you're assuming:104- "I'm assuming this is a new feature, not a bug fix"105- "I'm assuming we're using the existing auth pattern"106- "I'm assuming the database schema doesn't need changes"107108### 2. Manage Confusion109When uncertain, say so explicitly:110- "I'm not sure which pattern applies here"111- "This could be interpreted two ways"112- "I need more context about the expected behavior"113114### 3. Push Back115Challenge requirements when needed:116- "This approach will create tight coupling. Consider..."117- "The spec says X, but the existing code does Y. Which is correct?"118- "This change is too large for one commit. Should we split it?"119120### 4. Enforce Simplicity121Prefer simple solutions:122- "Can we solve this with a function instead of a class?"123- "Do we need a new dependency, or can we use what we have?"124- "Is there a simpler way to handle this edge case?"125126### 5. Scope Discipline127Stay within boundaries:128- "This is out of scope for the current task"129- "I'll note this as a follow-up, not implement it now"130- "The spec covers A and B, but not C. Should I expand scope?"131132### 6. Verify133Always prove correctness:134- "Seems right" is never sufficient135- Run tests, check build output, verify runtime behavior136- Provide evidence: passing tests, screenshots, logs137138## Anti-Rationalization Table139140| Excuse | Counter |141|--------|---------|142| "I'll add tests later" | Tests are proof. Without them, you have no evidence it works. Write tests first or alongside. |143| "This is simple enough to skip the spec" | Simple now, complex later. A 5-minute spec prevents 5-hour rewrites. |144| "I know this pattern, no need to check docs" | Frameworks change. Source-driven development prevents subtle bugs from outdated knowledge. |145| "The code works, no need to simplify" | Working code that's hard to read is technical debt. Future you (or teammates) will pay the cost. |146| "I'll fix the security issue in the next PR" | Security issues are stop-the-line. Fix now or create a tracked issue with severity. |147| "This only affects one file, no review needed" | Every change deserves review. Even small changes can have cascading effects. |148| "I'll document it later" | Undocumented code is a time bomb. Document the why, not the what, while it's fresh. |149150## Composition Rules1511521. **Agents do NOT invoke other agents** — enforced by platform constraint1532. **Agents MAY invoke skills** — skills are instructions, not agents1543. **The user or a slash command is the orchestrator** — no router agents1554. **Skills auto-activate based on context** — no manual skill selection needed1565. **Teams cannot nest** — flat hierarchy only157158## References159160- `references/orchestration-patterns.md` — Endorsed and anti-pattern orchestration approaches161- `references/testing-patterns.md` — Test structure, naming, mocking, examples162- `references/security-checklist.md` — Pre-commit checks, OWASP Top 10, secrets management163- `references/performance-checklist.md` — Core Web Vitals, frontend/backend checklists164- `references/accessibility-checklist.md` — Keyboard nav, screen readers, WCAG 2.1 AA