Speckit Checklist Skill
Checklist Purpose: "Unit Tests for English"
CRITICAL CONCEPT: Checklists are UNIT TESTS FOR REQUIREMENTS WRITING - they validate the quality, clarity, and completeness of requirements in a given domain.
NOT for verification/testing:
- ❌ NOT "Verify the button clicks correctly"
- ❌ NOT "Test error handling works"
- ❌ NOT "Confirm the API returns 200"
- ❌ NOT checking if code/implementation matches the spec
FOR requirements quality validation:
- ✅ "Are visual hierarchy requirements defined for all card types?" (completeness)
- ✅ "Is 'prominent display' quantified with specific sizing/positioning?" (clarity)
- ✅ "Are hover state requirements consistent across all interactive elements?" (consistency)
- ✅ "Are accessibility requirements defined for keyboard navigation?" (coverage)
- ✅ "Does the spec define what happens when logo image fails to load?" (edge cases)
Metaphor: If your spec is code written in English, the checklist is its unit test suite. You're testing whether the requirements are well-written, complete, unambiguous, and ready for implementation - NOT whether the implementation works.
User Provided Context
{{user_provided_context}}
You MUST consider the User Provided Context before proceeding (if not empty).
Execution Steps
Setup: Run .specify/scripts/bash/check-prerequisites.sh --json from repo root and parse JSON for FEATURE_DIR and AVAILABLE_DOCS list.
- All file paths must be absolute.
- For single quotes in args like "I'm Groot", use escape syntax: e.g 'I'''m Groot' (or double-quote if possible: "I'm Groot").
Clarify intent (dynamic): Derive up to THREE initial contextual clarifying questions (no pre-baked catalog). They MUST:
- Be generated from the user's phrasing + extracted signals from spec/plan/tasks
- Only ask about information that materially changes checklist content
- Be skipped individually if already unambiguous in
{{user_provided_context}}
- Prefer precision over breadth
Generation algorithm:
- Extract signals: feature domain keywords (e.g., auth, latency, UX, API), risk indicators ("critical", "must", "compliance"), stakeholder hints ("QA", "review", "security team"), and explicit deliverables ("a11y", "rollback", "contracts").
- Cluster signals into candidate focus areas (max 4) ranked by relevance.
- Identify probable audience & timing (author, reviewer, QA, release) if not explicit.
- Detect missing dimensions: scope breadth, depth/rigor, risk emphasis, exclusion boundaries, measurable acceptance criteria.
- Formulate questions chosen from these archetypes:
- Scope refinement (e.g., "Should this include integration touchpoints with X and Y or stay limited to local module correctness?")
- Risk prioritization (e.g., "Which of these potential risk areas should receive mandatory gating checks?")
- Depth calibration (e.g., "Is this a lightweight pre-commit sanity list or a formal release gate?")
- Audience framing (e.g., "Will this be used by the author only or peers during PR review?")
- Boundary exclusion (e.g., "Should we explicitly exclude performance tuning items this round?")
- Scenario class gap (e.g., "No recovery flows detected—are rollback / partial failure paths in scope?")
Question formatting rules:
- If presenting options, generate a compact table with columns: Option | Candidate | Why It Matters
- Limit to A–E options maximum; omit table if a free-form answer is clearer
- Never ask the user to restate what they already said
- Avoid speculative categories (no hallucination). If uncertain, ask explicitly: "Confirm whether X belongs in scope."
Defaults when interaction impossible:
- Depth: Standard
- Audience: Reviewer (PR) if code-related; Author otherwise
- Focus: Top 2 relevance clusters
Output the questions (label Q1/Q2/Q3). After answers: if ≥2 scenario classes (Alternate / Exception / Recovery / Non-Functional domain) remain unclear, you MAY ask up to TWO more targeted follow‑ups (Q4/Q5) with a one-line justification each (e.g., "Unresolved recovery path risk"). Do not exceed five total questions. Skip escalation if user explicitly declines more.
Understand user request: Combine {{user_provided_context}} + clarifying answers:
- Derive checklist theme (e.g., security, review, deploy, ux)
- Consolidate explicit must-have items mentioned by user
- Map focus selections to category scaffolding
- Infer any missing context from spec/plan/tasks (do NOT hallucinate)
Load feature context: Read from FEATURE_DIR:
- spec.md: Feature requirements and scope
- plan.md (if exists): Technical details, dependencies
- tasks.md (if exists): Implementation tasks
Context Loading Strategy:
- Load only necessary portions relevant to active focus areas (avoid full-file dumping)
- Prefer summarizing long sections into concise scenario/requirement bullets
- Use progressive disclosure: add follow-on retrieval only if gaps detected
- If source docs are large, generate interim summary items instead of embedding raw text
Generate checklist - Create "Unit Tests for Requirements":
- Create
FEATURE_DIR/checklists/ directory if it doesn't exist
- Generate unique checklist filename:
- Use short, descriptive name based on domain (e.g.,
ux.md, api.md, security.md)
- Format:
[domain].md
- If file exists, append to existing file
- Number items sequentially starting from CHK001
- Each
speckit-checklist run creates a NEW file (never overwrites existing checklists)
CORE PRINCIPLE - Test the Requirements, Not the Implementation:
Every checklist item MUST evaluate the REQUIREMENTS THEMSELVES for:
- Completeness: Are all necessary requirements present?
- Clarity: Are requirements unambiguous and specific?
- Consistency: Do requirements align with each other?
- Measurability: Can requirements be objectively verified?
- Coverage: Are all scenarios/edge cases addressed?
Category Structure - Group items by requirement quality dimensions:
- Requirement Completeness (Are all necessary requirements documented?)
- Requirement Clarity (Are requirements specific and unambiguous?)
- Requirement Consistency (Do requirements align without conflicts?)
- Acceptance Criteria Quality (Are success criteria measurable?)
- Scenario Coverage (Are all flows/cases addressed?)
- Edge Case Coverage (Are boundary conditions defined?)
- Non-Functional Requirements (Performance, Security, Accessibility, etc. - are they specified?)
- Dependencies & Assumptions (Are they documented and validated?)
- Ambiguities & Conflicts (What needs clarification?)
HOW TO WRITE CHECKLIST ITEMS - "Unit Tests for English":
❌ WRONG (Testing implementation):
- "Verify landing page displays 3 episode cards"
- "Test hover states work on desktop"
- "Confirm logo click navigates home"
✅ CORRECT (Testing requirements quality):
- "Are the exact number and layout of featured episodes specified?" [Completeness]
- "Is 'prominent display' quantified with specific sizing/positioning?" [Clarity]
- "Are hover state requirements consistent across all interactive elements?" [Consistency]
- "Are keyboard navigation requirements defined for all interactive UI?" [Coverage]
- "Is the fallback behavior specified when logo image fails to load?" [Edge Cases]
- "Are loading states defined for asynchronous episode data?" [Completeness]
- "Does the spec define visual hierarchy for competing UI elements?" [Clarity]
ITEM STRUCTURE:
Each item should follow this pattern:
- Question format asking about requirement quality
- Focus on what's WRITTEN (or not written) in the spec/plan
- Include quality dimension in brackets [Completeness/Clarity/Consistency/etc.]
- Reference spec section
[Spec §X.Y] when checking existing requirements
- Use
[Gap] marker when checking for missing requirements
EXAMPLES BY QUALITY DIMENSION:
Completeness:
- "Are error handling requirements defined for all API failure modes? [Gap]"
- "Are accessibility requirements specified for all interactive elements? [Completeness]"
- "Are mobile breakpoint requirements defined for responsive layouts? [Gap]"
Clarity:
- "Is 'fast loading' quantified with specific timing thresholds? [Clarity, Spec §NFR-2]"
- "Are 'related episodes' selection criteria explicitly defined? [Clarity, Spec §FR-5]"
- "Is 'prominent' defined with measurable visual properties? [Ambiguity, Spec §FR-4]"
Consistency:
- "Do navigation requirements align across all pages? [Consistency, Spec §FR-10]"
- "Are card component requirements consistent between landing and detail pages? [Consistency]"
Coverage:
- "Are requirements defined for zero-state scenarios (no episodes)? [Coverage, Edge Case]"
- "Are concurrent user interaction scenarios addressed? [Coverage, Gap]"
- "Are requirements specified for partial data loading failures? [Coverage, Exception Flow]"
Measurability:
- "Are visual hierarchy requirements measurable/testable? [Acceptance Criteria, Spec §FR-1]"
- "Can 'balanced visual weight' be objectively verified? [Measurability, Spec §FR-2]"
Scenario Classification & Coverage (Requirements Quality Focus):
- Check if requirements exist for: Primary, Alternate, Exception/Error, Recovery, Non-Functional scenarios
- For each scenario class, ask: "Are [scenario type] requirements complete, clear, and consistent?"
- If scenario class missing: "Are [scenario type] requirements intentionally excluded or missing? [Gap]"
- Include resilience/rollback when state mutation occurs: "Are rollback requirements defined for migration failures? [Gap]"
Traceability Requirements:
- MINIMUM: ≥80% of items MUST include at least one traceability reference
- Each item should reference: spec section
[Spec §X.Y], or use markers: [Gap], [Ambiguity], [Conflict], [Assumption]
- If no ID system exists: "Is a requirement & acceptance criteria ID scheme established? [Traceability]"
Surface & Resolve Issues (Requirements Quality Problems):
Ask questions about the requirements themselves:
- Ambiguities: "Is the term 'fast' quantified with specific metrics? [Ambiguity, Spec §NFR-1]"
- Conflicts: "Do navigation requirements conflict between §FR-10 and §FR-10a? [Conflict]"
- Assumptions: "Is the assumption of 'always available podcast API' validated? [Assumption]"
- Dependencies: "Are external podcast API requirements documented? [Dependency, Gap]"
- Missing definitions: "Is 'visual hierarchy' defined with measurable criteria? [Gap]"
Content Consolidation:
- Soft cap: If raw candidate items > 40, prioritize by risk/impact
- Merge near-duplicates checking the same requirement aspect
- If >5 low-impact edge cases, create one item: "Are edge cases X, Y, Z addressed in requirements? [Coverage]"
🚫 ABSOLUTELY PROHIBITED - These make it an implementation test, not a requirements test:
- ❌ Any item starting with "Verify", "Test", "Confirm", "Check" + implementation behavior
- ❌ References to code execution, user actions, system behavior
- ❌ "Displays correctly", "works properly", "functions as expected"
- ❌ "Click", "navigate", "render", "load", "execute"
- ❌ Test cases, test plans, QA procedures
- ❌ Implementation details (frameworks, APIs, algorithms)
✅ REQUIRED PATTERNS - These test requirements quality:
- ✅ "Are [requirement type] defined/specified/documented for [scenario]?"
- ✅ "Is [vague term] quantified/clarified with specific criteria?"
- ✅ "Are requirements consistent between [section A] and [section B]?"
- ✅ "Can [requirement] be objectively measured/verified?"
- ✅ "Are [edge cases/scenarios] addressed in requirements?"
- ✅ "Does the spec define [missing aspect]?"
Structure Reference: Generate the checklist following the canonical template in .specify/templates/checklist-template.md for title, meta section, category headings, and ID formatting. If template is unavailable, use: H1 title, purpose/created meta lines, ## category sections containing - [ ] CHK### <requirement item> lines with globally incrementing IDs starting at CHK001.
Report: Output full path to created checklist, item count, and remind user that each run creates a new file. Summarize:
- Focus areas selected
- Depth level
- Actor/timing
- Any explicit user-specified must-have items incorporated
Important: Each speckit-checklist skill invocation creates a checklist file using short, descriptive names unless file already exists. This allows:
- Multiple checklists of different types (e.g.,
ux.md, test.md, security.md)
- Simple, memorable filenames that indicate checklist purpose
- Easy identification and navigation in the
checklists/ folder
To avoid clutter, use descriptive types and clean up obsolete checklists when done.
Example Checklist Types & Sample Items
UX Requirements Quality: ux.md
Sample items (testing the requirements, NOT the implementation):
- "Are visual hierarchy requirements defined with measurable criteria? [Clarity, Spec §FR-1]"
- "Is the number and positioning of UI elements explicitly specified? [Completeness, Spec §FR-1]"
- "Are interaction state requirements (hover, focus, active) consistently defined? [Consistency]"
- "Are accessibility requirements specified for all interactive elements? [Coverage, Gap]"
- "Is fallback behavior defined when images fail to load? [Edge Case, Gap]"
- "Can 'prominent display' be objectively measured? [Measurability, Spec §FR-4]"
API Requirements Quality: api.md
Sample items:
- "Are error response formats specified for all failure scenarios? [Completeness]"
- "Are rate limiting requirements quantified with specific thresholds? [Clarity]"
- "Are authentication requirements consistent across all endpoints? [Consistency]"
- "Are retry/timeout requirements defined for external dependencies? [Coverage, Gap]"
- "Is versioning strategy documented in requirements? [Gap]"
Performance Requirements Quality: performance.md
Sample items:
- "Are performance requirements quantified with specific metrics? [Clarity]"
- "Are performance targets defined for all critical user journeys? [Coverage]"
- "Are performance requirements under different load conditions specified? [Completeness]"
- "Can performance requirements be objectively measured? [Measurability]"
- "Are degradation requirements defined for high-load scenarios? [Edge Case, Gap]"
Security Requirements Quality: security.md
Sample items:
- "Are authentication requirements specified for all protected resources? [Coverage]"
- "Are data protection requirements defined for sensitive information? [Completeness]"
- "Is the threat model documented and requirements aligned to it? [Traceability]"
- "Are security requirements consistent with compliance obligations? [Consistency]"
- "Are security failure/breach response requirements defined? [Gap, Exception Flow]"
Anti-Examples: What NOT To Do
❌ WRONG - These test implementation, not requirements:
- [ ] CHK001 - Verify landing page displays 3 episode cards [Spec §FR-001]
- [ ] CHK002 - Test hover states work correctly on desktop [Spec §FR-003]
- [ ] CHK003 - Confirm logo click navigates to home page [Spec §FR-010]
- [ ] CHK004 - Check that related episodes section shows 3-5 items [Spec §FR-005]
✅ CORRECT - These test requirements quality:
- [ ] CHK001 - Are the number and layout of featured episodes explicitly specified? [Completeness, Spec §FR-001]
- [ ] CHK002 - Are hover state requirements consistently defined for all interactive elements? [Consistency, Spec §FR-003]
- [ ] CHK003 - Are navigation requirements clear for all clickable brand elements? [Clarity, Spec §FR-010]
- [ ] CHK004 - Is the selection criteria for related episodes documented? [Gap, Spec §FR-005]
- [ ] CHK005 - Are loading state requirements defined for asynchronous episode data? [Gap]
- [ ] CHK006 - Can "visual hierarchy" requirements be objectively measured? [Measurability, Spec §FR-001]
Key Differences:
- Wrong: Tests if the system works correctly
- Correct: Tests if the requirements are written correctly
- Wrong: Verification of behavior
- Correct: Validation of requirement quality
- Wrong: "Does it do X?"
- Correct: "Is X clearly specified?"
1---2name: speckit-checklist3description: Generate quality checklists to validate requirements completeness, clarity, and consistency.4---5
6# Speckit Checklist Skill
7
8## Checklist Purpose: "Unit Tests for English"
9
10**CRITICAL CONCEPT**: Checklists are **UNIT TESTS FOR REQUIREMENTS WRITING** - they validate the quality, clarity, and completeness of requirements in a given domain.
11
12**NOT for verification/testing**:
13
14- ❌ NOT "Verify the button clicks correctly"
15- ❌ NOT "Test error handling works"
16- ❌ NOT "Confirm the API returns 200"
17- ❌ NOT checking if code/implementation matches the spec
18
19**FOR requirements quality validation**:
20
21- ✅ "Are visual hierarchy requirements defined for all card types?" (completeness)
22- ✅ "Is 'prominent display' quantified with specific sizing/positioning?" (clarity)
23- ✅ "Are hover state requirements consistent across all interactive elements?" (consistency)
24- ✅ "Are accessibility requirements defined for keyboard navigation?" (coverage)
25- ✅ "Does the spec define what happens when logo image fails to load?" (edge cases)
26
27**Metaphor**: If your spec is code written in English, the checklist is its unit test suite. You're testing whether the requirements are well-written, complete, unambiguous, and ready for implementation - NOT whether the implementation works.
28
29## User Provided Context
30
31```text
32{{user_provided_context}}
33```
34
35You **MUST** consider the User Provided Context before proceeding (if not empty).
36
37## Execution Steps
38
391. **Setup**: Run `.specify/scripts/bash/check-prerequisites.sh --json` from repo root and parse JSON for FEATURE_DIR and AVAILABLE_DOCS list.
40 - All file paths must be absolute.
41 - For single quotes in args like "I'm Groot", use escape syntax: e.g 'I'''m Groot' (or double-quote if possible: "I'm Groot").
42
432. **Clarify intent (dynamic)**: Derive up to THREE initial contextual clarifying questions (no pre-baked catalog). They MUST:
44 - Be generated from the user's phrasing + extracted signals from spec/plan/tasks
45 - Only ask about information that materially changes checklist content
46 - Be skipped individually if already unambiguous in `{{user_provided_context}}`
47 - Prefer precision over breadth
48
49 Generation algorithm:
50 1. Extract signals: feature domain keywords (e.g., auth, latency, UX, API), risk indicators ("critical", "must", "compliance"), stakeholder hints ("QA", "review", "security team"), and explicit deliverables ("a11y", "rollback", "contracts").
51 2. Cluster signals into candidate focus areas (max 4) ranked by relevance.
52 3. Identify probable audience & timing (author, reviewer, QA, release) if not explicit.
53 4. Detect missing dimensions: scope breadth, depth/rigor, risk emphasis, exclusion boundaries, measurable acceptance criteria.
54 5. Formulate questions chosen from these archetypes:
55 - Scope refinement (e.g., "Should this include integration touchpoints with X and Y or stay limited to local module correctness?")
56 - Risk prioritization (e.g., "Which of these potential risk areas should receive mandatory gating checks?")
57 - Depth calibration (e.g., "Is this a lightweight pre-commit sanity list or a formal release gate?")
58 - Audience framing (e.g., "Will this be used by the author only or peers during PR review?")
59 - Boundary exclusion (e.g., "Should we explicitly exclude performance tuning items this round?")
60 - Scenario class gap (e.g., "No recovery flows detected—are rollback / partial failure paths in scope?")
61
62 Question formatting rules:
63 - If presenting options, generate a compact table with columns: Option | Candidate | Why It Matters
64 - Limit to A–E options maximum; omit table if a free-form answer is clearer
65 - Never ask the user to restate what they already said
66 - Avoid speculative categories (no hallucination). If uncertain, ask explicitly: "Confirm whether X belongs in scope."
67
68 Defaults when interaction impossible:
69 - Depth: Standard
70 - Audience: Reviewer (PR) if code-related; Author otherwise
71 - Focus: Top 2 relevance clusters
72
73 Output the questions (label Q1/Q2/Q3). After answers: if ≥2 scenario classes (Alternate / Exception / Recovery / Non-Functional domain) remain unclear, you MAY ask up to TWO more targeted follow‑ups (Q4/Q5) with a one-line justification each (e.g., "Unresolved recovery path risk"). Do not exceed five total questions. Skip escalation if user explicitly declines more.
74
753. **Understand user request**: Combine `{{user_provided_context}}` + clarifying answers:
76 - Derive checklist theme (e.g., security, review, deploy, ux)
77 - Consolidate explicit must-have items mentioned by user
78 - Map focus selections to category scaffolding
79 - Infer any missing context from spec/plan/tasks (do NOT hallucinate)
80
814. **Load feature context**: Read from FEATURE_DIR:
82 - spec.md: Feature requirements and scope
83 - plan.md (if exists): Technical details, dependencies
84 - tasks.md (if exists): Implementation tasks
85
86 **Context Loading Strategy**:
87 - Load only necessary portions relevant to active focus areas (avoid full-file dumping)
88 - Prefer summarizing long sections into concise scenario/requirement bullets
89 - Use progressive disclosure: add follow-on retrieval only if gaps detected
90 - If source docs are large, generate interim summary items instead of embedding raw text
91
925. **Generate checklist** - Create "Unit Tests for Requirements":
93 - Create `FEATURE_DIR/checklists/` directory if it doesn't exist
94 - Generate unique checklist filename:
95 - Use short, descriptive name based on domain (e.g., `ux.md`, `api.md`, `security.md`)
96 - Format: `[domain].md`
97 - If file exists, append to existing file
98 - Number items sequentially starting from CHK001
99 - Each `speckit-checklist` run creates a NEW file (never overwrites existing checklists)
100
101 **CORE PRINCIPLE - Test the Requirements, Not the Implementation**:
102 Every checklist item MUST evaluate the REQUIREMENTS THEMSELVES for:
103 - **Completeness**: Are all necessary requirements present?
104 - **Clarity**: Are requirements unambiguous and specific?
105 - **Consistency**: Do requirements align with each other?
106 - **Measurability**: Can requirements be objectively verified?
107 - **Coverage**: Are all scenarios/edge cases addressed?
108
109 **Category Structure** - Group items by requirement quality dimensions:
110 - **Requirement Completeness** (Are all necessary requirements documented?)
111 - **Requirement Clarity** (Are requirements specific and unambiguous?)
112 - **Requirement Consistency** (Do requirements align without conflicts?)
113 - **Acceptance Criteria Quality** (Are success criteria measurable?)
114 - **Scenario Coverage** (Are all flows/cases addressed?)
115 - **Edge Case Coverage** (Are boundary conditions defined?)
116 - **Non-Functional Requirements** (Performance, Security, Accessibility, etc. - are they specified?)
117 - **Dependencies & Assumptions** (Are they documented and validated?)
118 - **Ambiguities & Conflicts** (What needs clarification?)
119
120 **HOW TO WRITE CHECKLIST ITEMS - "Unit Tests for English"**:
121
122 ❌ **WRONG** (Testing implementation):
123 - "Verify landing page displays 3 episode cards"
124 - "Test hover states work on desktop"
125 - "Confirm logo click navigates home"
126
127 ✅ **CORRECT** (Testing requirements quality):
128 - "Are the exact number and layout of featured episodes specified?" [Completeness]
129 - "Is 'prominent display' quantified with specific sizing/positioning?" [Clarity]
130 - "Are hover state requirements consistent across all interactive elements?" [Consistency]
131 - "Are keyboard navigation requirements defined for all interactive UI?" [Coverage]
132 - "Is the fallback behavior specified when logo image fails to load?" [Edge Cases]
133 - "Are loading states defined for asynchronous episode data?" [Completeness]
134 - "Does the spec define visual hierarchy for competing UI elements?" [Clarity]
135
136 **ITEM STRUCTURE**:
137 Each item should follow this pattern:
138 - Question format asking about requirement quality
139 - Focus on what's WRITTEN (or not written) in the spec/plan
140 - Include quality dimension in brackets [Completeness/Clarity/Consistency/etc.]
141 - Reference spec section `[Spec §X.Y]` when checking existing requirements
142 - Use `[Gap]` marker when checking for missing requirements
143
144 **EXAMPLES BY QUALITY DIMENSION**:
145
146 Completeness:
147 - "Are error handling requirements defined for all API failure modes? [Gap]"
148 - "Are accessibility requirements specified for all interactive elements? [Completeness]"
149 - "Are mobile breakpoint requirements defined for responsive layouts? [Gap]"
150
151 Clarity:
152 - "Is 'fast loading' quantified with specific timing thresholds? [Clarity, Spec §NFR-2]"
153 - "Are 'related episodes' selection criteria explicitly defined? [Clarity, Spec §FR-5]"
154 - "Is 'prominent' defined with measurable visual properties? [Ambiguity, Spec §FR-4]"
155
156 Consistency:
157 - "Do navigation requirements align across all pages? [Consistency, Spec §FR-10]"
158 - "Are card component requirements consistent between landing and detail pages? [Consistency]"
159
160 Coverage:
161 - "Are requirements defined for zero-state scenarios (no episodes)? [Coverage, Edge Case]"
162 - "Are concurrent user interaction scenarios addressed? [Coverage, Gap]"
163 - "Are requirements specified for partial data loading failures? [Coverage, Exception Flow]"
164
165 Measurability:
166 - "Are visual hierarchy requirements measurable/testable? [Acceptance Criteria, Spec §FR-1]"
167 - "Can 'balanced visual weight' be objectively verified? [Measurability, Spec §FR-2]"
168
169 **Scenario Classification & Coverage** (Requirements Quality Focus):
170 - Check if requirements exist for: Primary, Alternate, Exception/Error, Recovery, Non-Functional scenarios
171 - For each scenario class, ask: "Are [scenario type] requirements complete, clear, and consistent?"
172 - If scenario class missing: "Are [scenario type] requirements intentionally excluded or missing? [Gap]"
173 - Include resilience/rollback when state mutation occurs: "Are rollback requirements defined for migration failures? [Gap]"
174
175 **Traceability Requirements**:
176 - MINIMUM: ≥80% of items MUST include at least one traceability reference
177 - Each item should reference: spec section `[Spec §X.Y]`, or use markers: `[Gap]`, `[Ambiguity]`, `[Conflict]`, `[Assumption]`
178 - If no ID system exists: "Is a requirement & acceptance criteria ID scheme established? [Traceability]"
179
180 **Surface & Resolve Issues** (Requirements Quality Problems):
181 Ask questions about the requirements themselves:
182 - Ambiguities: "Is the term 'fast' quantified with specific metrics? [Ambiguity, Spec §NFR-1]"
183 - Conflicts: "Do navigation requirements conflict between §FR-10 and §FR-10a? [Conflict]"
184 - Assumptions: "Is the assumption of 'always available podcast API' validated? [Assumption]"
185 - Dependencies: "Are external podcast API requirements documented? [Dependency, Gap]"
186 - Missing definitions: "Is 'visual hierarchy' defined with measurable criteria? [Gap]"
187
188 **Content Consolidation**:
189 - Soft cap: If raw candidate items > 40, prioritize by risk/impact
190 - Merge near-duplicates checking the same requirement aspect
191 - If >5 low-impact edge cases, create one item: "Are edge cases X, Y, Z addressed in requirements? [Coverage]"
192
193 **🚫 ABSOLUTELY PROHIBITED** - These make it an implementation test, not a requirements test:
194 - ❌ Any item starting with "Verify", "Test", "Confirm", "Check" + implementation behavior
195 - ❌ References to code execution, user actions, system behavior
196 - ❌ "Displays correctly", "works properly", "functions as expected"
197 - ❌ "Click", "navigate", "render", "load", "execute"
198 - ❌ Test cases, test plans, QA procedures
199 - ❌ Implementation details (frameworks, APIs, algorithms)
200
201 **✅ REQUIRED PATTERNS** - These test requirements quality:
202 - ✅ "Are [requirement type] defined/specified/documented for [scenario]?"
203 - ✅ "Is [vague term] quantified/clarified with specific criteria?"
204 - ✅ "Are requirements consistent between [section A] and [section B]?"
205 - ✅ "Can [requirement] be objectively measured/verified?"
206 - ✅ "Are [edge cases/scenarios] addressed in requirements?"
207 - ✅ "Does the spec define [missing aspect]?"
208
2096. **Structure Reference**: Generate the checklist following the canonical template in `.specify/templates/checklist-template.md` for title, meta section, category headings, and ID formatting. If template is unavailable, use: H1 title, purpose/created meta lines, `##` category sections containing `- [ ] CHK### <requirement item>` lines with globally incrementing IDs starting at CHK001.
210
2117. **Report**: Output full path to created checklist, item count, and remind user that each run creates a new file. Summarize:
212 - Focus areas selected
213 - Depth level
214 - Actor/timing
215 - Any explicit user-specified must-have items incorporated
216
217**Important**: Each `speckit-checklist` skill invocation creates a checklist file using short, descriptive names unless file already exists. This allows:
218
219- Multiple checklists of different types (e.g., `ux.md`, `test.md`, `security.md`)
220- Simple, memorable filenames that indicate checklist purpose
221- Easy identification and navigation in the `checklists/` folder
222
223To avoid clutter, use descriptive types and clean up obsolete checklists when done.
224
225## Example Checklist Types & Sample Items
226
227**UX Requirements Quality:** `ux.md`
228
229Sample items (testing the requirements, NOT the implementation):
230
231- "Are visual hierarchy requirements defined with measurable criteria? [Clarity, Spec §FR-1]"
232- "Is the number and positioning of UI elements explicitly specified? [Completeness, Spec §FR-1]"
233- "Are interaction state requirements (hover, focus, active) consistently defined? [Consistency]"
234- "Are accessibility requirements specified for all interactive elements? [Coverage, Gap]"
235- "Is fallback behavior defined when images fail to load? [Edge Case, Gap]"
236- "Can 'prominent display' be objectively measured? [Measurability, Spec §FR-4]"
237
238**API Requirements Quality:** `api.md`
239
240Sample items:
241
242- "Are error response formats specified for all failure scenarios? [Completeness]"
243- "Are rate limiting requirements quantified with specific thresholds? [Clarity]"
244- "Are authentication requirements consistent across all endpoints? [Consistency]"
245- "Are retry/timeout requirements defined for external dependencies? [Coverage, Gap]"
246- "Is versioning strategy documented in requirements? [Gap]"
247
248**Performance Requirements Quality:** `performance.md`
249
250Sample items:
251
252- "Are performance requirements quantified with specific metrics? [Clarity]"
253- "Are performance targets defined for all critical user journeys? [Coverage]"
254- "Are performance requirements under different load conditions specified? [Completeness]"
255- "Can performance requirements be objectively measured? [Measurability]"
256- "Are degradation requirements defined for high-load scenarios? [Edge Case, Gap]"
257
258**Security Requirements Quality:** `security.md`
259
260Sample items:
261
262- "Are authentication requirements specified for all protected resources? [Coverage]"
263- "Are data protection requirements defined for sensitive information? [Completeness]"
264- "Is the threat model documented and requirements aligned to it? [Traceability]"
265- "Are security requirements consistent with compliance obligations? [Consistency]"
266- "Are security failure/breach response requirements defined? [Gap, Exception Flow]"
267
268## Anti-Examples: What NOT To Do
269
270**❌ WRONG - These test implementation, not requirements:**
271
272```markdown
273- [ ] CHK001 - Verify landing page displays 3 episode cards [Spec §FR-001]
274- [ ] CHK002 - Test hover states work correctly on desktop [Spec §FR-003]
275- [ ] CHK003 - Confirm logo click navigates to home page [Spec §FR-010]
276- [ ] CHK004 - Check that related episodes section shows 3-5 items [Spec §FR-005]
277```
278
279**✅ CORRECT - These test requirements quality:**
280
281```markdown
282- [ ] CHK001 - Are the number and layout of featured episodes explicitly specified? [Completeness, Spec §FR-001]
283- [ ] CHK002 - Are hover state requirements consistently defined for all interactive elements? [Consistency, Spec §FR-003]
284- [ ] CHK003 - Are navigation requirements clear for all clickable brand elements? [Clarity, Spec §FR-010]
285- [ ] CHK004 - Is the selection criteria for related episodes documented? [Gap, Spec §FR-005]
286- [ ] CHK005 - Are loading state requirements defined for asynchronous episode data? [Gap]
287- [ ] CHK006 - Can "visual hierarchy" requirements be objectively measured? [Measurability, Spec §FR-001]
288```
289
290**Key Differences:**
291
292- Wrong: Tests if the system works correctly
293- Correct: Tests if the requirements are written correctly
294- Wrong: Verification of behavior
295- Correct: Validation of requirement quality
296- Wrong: "Does it do X?"
297- Correct: "Is X clearly specified?"