Talk Stage 1: Extract
Transforms raw material (article, transcript, notes, or a mix) into a structured summary ready for the pipeline's downstream stages. Auto-detects source type.
When to Use This Skill
- Starting a new talk from any source material
- First step of the talk pipeline (always run before other stages)
- Auditing existing source material before committing to a talk
What This Skill Does
- Collects metadata by asking for slug, event, date, duration, audience, mode if not provided
- Reads the source and loads the source file or inline content
- Detects source type (REX with real-world proof vs Concept with ideas/thesis) based on content signals
- Extracts the narrative arc (chronological for REX, thematic for Concept)
- Extracts metrics with every measurable number and its source
- Identifies main themes (3-7 themes)
- Flags gaps: what's missing for a complete talk
- Writes
{slug}-summary.md
Input
Required:
- Source file path or inline content (article
.mdx, transcript .md, notes)
- Metadata:
slug, event, date, duration, audience, type (--rex or --concept)
If metadata is missing -> AskUserQuestion before proceeding.
Output
talks/{YYYY}-{slug}-summary.md
Source Type Detection
| REX signals |
Concept signals |
| Specific dates |
Theses, arguments |
| Measured metrics |
General observations |
| Project/tool names |
Trend observations |
| Commits, releases, PRs |
Analogies, metaphors |
| "I shipped", "We built" |
"I think", "In my opinion" |
If hybrid -> note both components in the summary.
Output Format
# Talk Summary: {Provisional Title}
**Slug** : {slug}
**Event** : {event}
**Date** : {date}
**Duration** : {duration} min
**Audience** : {audience description}
**Type detected** : REX | Concept | Hybrid
**Source** : {source file path}
---
## Narrative Arc
{Arc description: 3-5 sentences. Chronological if REX, thematic if Concept.}
## Main Themes
| # | Theme | Short description | Weight |
|---|-------|------------------|--------|
| 1 | {theme} | {description} | High/Medium/Low |
...
## Key Metrics Extracted
{All measurable numbers found in the source}
Format: `{value}`: {context} (Source: {section/page/git})
Examples:
- `1,200 commits` over 7 months (Source: "acceleration" section)
- `-97% traffic` after SSE migration (Source: CHANGELOG v1.1.0)
If none -> "No verifiable metrics found (Concept mode)"
## Narrative Potential
{3-5 sentences on the strengths and possible narrative angles.
What makes this talk potentially strong. What might be missing.}
## Gaps Identified
- [ ] {gap 1}: {how to fill it}
- [ ] {gap 2}: {how to fill it}
If no obvious gaps -> "No major gaps identified."
## Recommendations for next stages
- **Research**: {recommended / not applicable (Concept mode)}: {why}
- **Concepts**: {priority themes to explore}
- **Position**: {angles already visible from the source material}
---
*Generated by talk-stage1-extract: {date}*
*Source: {source path}*
Metric Extraction Rules
- Do not round without indicating it
- Always include the metric's source
- If two sources contradict, flag both without picking one
- No invented metrics to fill gaps
- Use
{before} -> {after} format for evolutions
Anti-patterns
- Vague summary ("This text is about AI...")
- Omitting metrics, even approximate ones with their source
- Hiding gaps (naming them is better than pretending they don't exist)
- Changing the detected type without justification
- Inventing a narrative arc not present in the source
Validation Checklist
Tips
- Run this before the orchestrator if you want to verify the source material is usable
- The summary is the foundation: every downstream stage reads it
- Hybrid sources (part REX, part Concept) are fine, name both components clearly
Related
1---2name: talk-stage1-extract3description: Extracts and structures source material (articles, transcripts, notes) into a talk summary with narrative arc, themes, metrics, and gaps. Auto-detects REX vs Concept type. Use when starting a new talk from any source material or auditing existing material before committing to a talk.4---5
6# Talk Stage 1: Extract
7
8Transforms raw material (article, transcript, notes, or a mix) into a structured summary ready for the pipeline's downstream stages. Auto-detects source type.
9
10## When to Use This Skill
11
12- Starting a new talk from any source material
13- First step of the talk pipeline (always run before other stages)
14- Auditing existing source material before committing to a talk
15
16## What This Skill Does
17
181. **Collects metadata** by asking for slug, event, date, duration, audience, mode if not provided
192. **Reads the source** and loads the source file or inline content
203. **Detects source type** (REX with real-world proof vs Concept with ideas/thesis) based on content signals
214. **Extracts the narrative arc** (chronological for REX, thematic for Concept)
225. **Extracts metrics** with every measurable number and its source
236. **Identifies main themes** (3-7 themes)
247. **Flags gaps**: what's missing for a complete talk
258. **Writes `{slug}-summary.md`**
26
27## Input
28
29Required:
30- Source file path or inline content (article `.mdx`, transcript `.md`, notes)
31- Metadata: `slug`, `event`, `date`, `duration`, `audience`, `type` (--rex or --concept)
32
33If metadata is missing -> `AskUserQuestion` before proceeding.
34
35## Output
36
37`talks/{YYYY}-{slug}-summary.md`
38
39## Source Type Detection
40
41| REX signals | Concept signals |
42|-------------|-----------------|
43| Specific dates | Theses, arguments |
44| Measured metrics | General observations |
45| Project/tool names | Trend observations |
46| Commits, releases, PRs | Analogies, metaphors |
47| "I shipped", "We built" | "I think", "In my opinion" |
48
49If hybrid -> note both components in the summary.
50
51## Output Format
52
53```markdown
54# Talk Summary: {Provisional Title}
55
56**Slug** : {slug}
57**Event** : {event}
58**Date** : {date}
59**Duration** : {duration} min
60**Audience** : {audience description}
61**Type detected** : REX | Concept | Hybrid
62**Source** : {source file path}
63
64---
65
66## Narrative Arc
67
68{Arc description: 3-5 sentences. Chronological if REX, thematic if Concept.}
69
70## Main Themes
71
72| # | Theme | Short description | Weight |
73|---|-------|------------------|--------|
74| 1 | {theme} | {description} | High/Medium/Low |
75...
76
77## Key Metrics Extracted
78
79{All measurable numbers found in the source}
80
81Format: `{value}`: {context} (Source: {section/page/git})
82
83Examples:
84- `1,200 commits` over 7 months (Source: "acceleration" section)
85- `-97% traffic` after SSE migration (Source: CHANGELOG v1.1.0)
86
87If none -> "No verifiable metrics found (Concept mode)"
88
89## Narrative Potential
90
91{3-5 sentences on the strengths and possible narrative angles.
92What makes this talk potentially strong. What might be missing.}
93
94## Gaps Identified
95
96- [ ] {gap 1}: {how to fill it}
97- [ ] {gap 2}: {how to fill it}
98
99If no obvious gaps -> "No major gaps identified."
100
101## Recommendations for next stages
102
103- **Research**: {recommended / not applicable (Concept mode)}: {why}
104- **Concepts**: {priority themes to explore}
105- **Position**: {angles already visible from the source material}
106
107---
108
109*Generated by talk-stage1-extract: {date}*
110*Source: {source path}*
111```
112
113## Metric Extraction Rules
114
115- Do not round without indicating it
116- Always include the metric's source
117- If two sources contradict, flag both without picking one
118- No invented metrics to fill gaps
119- Use `{before} -> {after}` format for evolutions
120
121## Anti-patterns
122
123- Vague summary ("This text is about AI...")
124- Omitting metrics, even approximate ones with their source
125- Hiding gaps (naming them is better than pretending they don't exist)
126- Changing the detected type without justification
127- Inventing a narrative arc not present in the source
128
129## Validation Checklist
130
131- [ ] Source type detected and justified
132- [ ] Narrative arc in 3-5 clear sentences
133- [ ] All measurable metrics extracted with their source
134- [ ] Main themes listed (3-7 max)
135- [ ] Gaps explicitly identified
136- [ ] File saved to `talks/{YYYY}-{slug}-summary.md`
137
138## Tips
139
140- Run this before the orchestrator if you want to verify the source material is usable
141- The summary is the foundation: every downstream stage reads it
142- Hybrid sources (part REX, part Concept) are fine, name both components clearly
143
144## Related
145
146- [Stage 2: Research](../stage-2-research/SKILL.md): git archaeology (REX mode)
147- [Stage 3: Concepts](../stage-3-concepts/SKILL.md): reads this summary
148- [Orchestrator](../orchestrator/SKILL.md): runs all stages in sequence