Interview Report: What You Know, What You Think
A round of interviews ends and the learning scatters: a hypothesis file
only its author can decode, debriefs nobody rereads, teammates asking
"so what did we actually find?" This skill closes the method by
distilling all of it into one report — what is now known (including
what got disproved), what is merely thought, and what remains
untested — every claim carrying its evidence, in the customers' own
words, organized so the people doing positioning, ideal-customer,
pricing, marketing, and product work can act on it.
The mental model
The report is the bridge from evidence to action
The interview method runs goals → hypotheses → questions → interviews →
learning. Its output is validated facts: hypotheses confirmed,
overturned, or tuned by real customer voices. Those facts are the raw
material of strategy — but there is no mechanical procedure that turns
"what customers said" into "what to do next." Humans do that combining,
and they can only do it if the facts arrive organized, honest about
their strength, and traceable to their sources. That package is this
report. It is deliberately NOT the strategy itself: it delivers the
evidence and names the decisions the evidence raises, and stops there.
Two readers, one document
The report serves both at once:
- The human skimmer reads only the top. So the report opens with a
summary that is as brief as possible without losing anything salient
— and each line is exactly three things: the status mark, the crisp
claim, the F-number. Nothing else. No citations or quotes; no
commentary or interpretation; no history of the finding — no "which
cuts against what we assumed," no "unlike our original hypothesis,"
no "ambiguous between X and Y." The status mark IS the entire
confidence-and-history a summary line gets; how the finding evolved,
what it contradicts, and what it might mean all live in the body.
Numbers may be part of the claim ("…one lost job ($300–800)");
explanations may not. If a summary line grows a "which…" clause or
an em-dash explanation, cut the clause and put it in the finding.
- The deep reader — a teammate doing the positioning work, or an
LLM assisting any downstream exercise, which reads everything
regardless of length — gets the reference sections below: every
finding with its voice-count, its debrief citations, and verbatim
quotes. In the body, ALWAYS quote and ALWAYS cite, with multiple
examples when multiple exist: the citations are simultaneously the
proof that a claim is correct and the trail for finding out more.
The finding-numbers (F1, F2, …) are the hinge between the two: a
summary line ends with its F-number, and the F-entry below carries the
evidence. The rule of thumb: when brevity is the goal, examples are
bloat; everywhere else, examples are the proof — they make a claim
believable, hard to counter, and easy for the reader to research
further.
Epistemic honesty is the product
The single way this report fails is by stating things more strongly
than the evidence supports — then every decision built on it inherits
the inflation. So every finding wears its status:
- ✓ Validated — a clear pattern across several conversations
confirmed it. Knowledge.
- ✗ Disproved — a clear pattern overturned it. Also knowledge —
negative knowledge is often the most valuable kind ("customers will
NOT pay extra for security" redirects an entire roadmap), so
disproved beliefs are stated prominently as things now known, never
buried as embarrassments.
- ~ Directional — supported but thin: few voices, or a single
revelation that reframed thinking but hasn't been re-tested. Stated
with its depth ("one voice, marked a revelation") and with what
would settle it.
- 👀 Watch — heard once or twice, noticed, parked. Imported from
the hypothesis file's watch section ("That's funny" items) and from
lone debrief observations.
- ? Untested — a goal, hypothesis, or question the interviews
never actually resolved: never asked, question misfired, segment
never reached. Named plainly, because knowing what you don't know is
what keeps downstream work honest.
Two more honesty obligations: sampling bias (who the interviewees
were, how they were recruited, who's missing — "all referrals, no
churned customers" changes what the findings mean) and "no pattern"
findings (when answers genuinely scattered, that's a finding —
knowing a pattern doesn't exist prevents building on a false one).
And state strength in numbers, never probability words. "Probably,"
"likely," "most," "often" mean wildly different things to different
readers — the differences between individual interpretations are larger
than the differences between the words — while "7 of 9" means the same
thing to everyone. Every tally is exact; if a directional finding needs
a confidence, be brave and put a number on it.
Splits are choices, never averages
When the evidence splits — five voices at $40, five at $300 — the
report never averages it into a mid-point nobody asked for. "It's a
balance" is usually a refusal to decide wearing the costume of
moderation. A split is either an emergent segment (report both
claims, each naming whom it's about, plus the markers that sort them)
or a genuine no-pattern finding (reported as such); the downstream
decision it raises is a choice, and the report names that choice
without making it. Related caution for readers, worth a line in the
product brief when relevant: tallies establish facts — where an
objectively true answer exists and individual errors cancel out — but
votes cannot design; averaging contradictory desires produces the
bland thing nobody hated enough to veto, not the thing anyone loves.
What the downstream work needs
Each per-area brief marshals findings for a specific exercise, in that
exercise's own terms:
- Ideal customer: keystone candidates (what the best interviewees
valued extremely — enough to drive a purchase by itself),
deal-breaker candidates (what disqualified the product regardless of
fit), the inciting-event stories actually heard (the specific
trigger moments that turned someone into an active buyer — quoted,
because these are collected precisely by interviewing), sorting
markers (behavioral and attitudinal characteristics that separate
best-fit from poor-fit — never mere demographics), and best-vs-worst
differentials.
- Positioning & messaging: the customers' exact vocabulary (what
they call themselves, the problem, the product category — their
words are the raw material of copy), the higher-level outcome they
are actually buying (what they said the product is for, one level
above what it does), the alternatives they compare against
(including do-it-yourself coping), and the vivid specifics — real
numbers, real emotions, real events — that make claims land.
- Pricing & packaging: willingness-to-pay bands with segment
attached, the anchors customers reason from ("that's what the
text-blast services cost"), budget and approval mechanics (whose
money, who signs, what threshold).
- Marketing & sales: where these customers discover and buy, whom
they trust, buyer vs. user vs. approver, the qualifying and
disqualifying signals heard.
- Product priorities: pains ranked by evidence (how many voices,
how hot the language), how customers cope today, unprompted feature
pulls, and what moved willingness-to-pay when mentioned.
Vocabulary
- Finding (F1, F2, …) — one distilled claim with a status mark,
evidence tally, debrief citations, and [H]/[G] references (an
emergent finding that maps to no hypothesis or goal carries no tag).
Numbers freeze when the report is finalized.
- Summary — the top section; salient claims only, F-numbered,
citation-free.
- Vocabulary bank — the customers' verbatim words and loaded
phrases, attributed.
- 💡 Tentative implication — the reporter's own labeled read,
allowed only in the per-area briefs, always marked as interpretation
and never presented as a finding.
- Decisions this raises — the stronger move: naming the choice the
evidence forces (e.g. "two segments — serving one is a strategy
decision") without making it.
The reporter's posture
Be clear, not clever
Write to be understood, not admired. The work here wrestles with hard
concepts, and clever metaphors, wordplay, or cute turns of phrase make
them harder to grasp, not easier. Say plainly what you mean. If a
sentence reads more clearly without a flourish, cut the flourish. State
the actual point rather than gesturing wittily at it.
Restate references; never cite a bare token
When you mention a numbered or lettered item to the user — K4, W2,
O17, H3, and the like — add a few plain words on what it actually is
("K4 — the owner whose career rides on the site"). A bare token is
unreadable to a human who saw it defined hours or days ago: the tag is
for traceability, the gloss is for comprehension. Keep the tag for
accuracy; always add the gloss.
Evidence or it doesn't get stated
Every finding in the body carries its tally ("7 of 9"), its debrief
citations by filename, and verbatim quotes — multiple examples whenever
multiple exist, because stacked independent voices are the proof. A
claim that can't cite a debrief doesn't go in the report. Never pad and
never trim: if only two voices support something, the tally says two
and the status says directional. Tallies name their denominator
honestly — "6 of 7 asked" when some debriefs lack the question — and a
voice never asked counts toward nothing, neither a claim nor its
disproof. Market-guru material (interviewees speaking for "most
people" rather than themselves) is excluded from tallies and kept in
the evidence with its flag. And tallies, like statuses, move only when
evidence is re-examined or added — never by rounding.
Status matches evidence — non-negotiable
The status marks are craft-gated: a ~ cannot be promoted to ✓ because
the user is confident, an ✗ cannot be softened because it's
disappointing, and an inconvenient finding cannot be dropped —
"make it look more validated for the investor deck" is refused, gently
and completely, because a report that flatters poisons everything
downstream and its readers can't tell. What the user rightfully owns:
wording (clearer phrasing of the same claim), emphasis (what the
summary leads with), scope (a section they'd rather omit — noted as
omitted), and their own interpretations, which are welcome in the
briefs when labeled as theirs. Record the user's judgment as judgment,
never as evidence.
Three refinements. Within an emergent segment, the denominator is the
segment — "5/5 multi-provider" can validate a segment-scoped claim. A
tally the sampling itself manufactured (9/9 name Facebook groups —
when all nine were recruited through one) caps the status at
directional no matter the count. And when the user disagrees with a
status and the evidence doesn't move, their position is recorded as a
plain labeled parenthetical beside the evidence inside the finding
("user's judgment, not evidence: expects this to validate"); the 💡
mark stays reserved for the briefs.
Brief on top, proof below
The summary contains no citations, no quotes, no commentary, and no
evolution story — brevity is its job, and every line is a bare claim
with its status mark, ending in the F-number that leads to the proof.
Even a disproved belief is stated as present-tense knowledge ("✗
Security does not drive buying (F2)"), not as narrative ("✗ our
security hypothesis was overturned"). The body contains ALL the
citations, quotes, comparisons, and explanation — completeness is its
job. Never blur the two: a summary that cites or explains is too long;
a body claim that doesn't cite is an opinion.
Interpretation is labeled
The skill may offer its own reads — "💡 this pattern smells like the
freelancer segment is the head of the market" — only inside the
per-area briefs, only marked 💡, and only phrased as interpretation.
Findings and implications never mix. When the evidence forces a choice
rather than suggesting an answer, prefer the "decisions this raises"
form and leave the choice with the humans.
The report reports
No positioning statements, no ideal-customer definitions, no price
recommendations, no roadmaps — the report feeds those exercises; it
does not preempt them. When the user asks for them ("so what should
our homepage say?"), point at the relevant brief and decline the rest:
that work deserves its own session with the report as input.
How to use this skill
Phase A — Ingest
Read what exists, asking only for what's missing:
- The working files: the hypothesis list (with its watch section
and change log), the question list, and the goal file. Accept any
subset — a missing goal file just means findings carry only [H]
references — but say what's absent and what the report loses.
- The debriefs: the directory of per-interview files. These are
the primary sources every finding will cite. Raw transcripts or
loose notes offered instead should be put on the record first — a
brief per-conversation file, answers mapped to questions — via a
debrief-recording skill such as Interview Debrief /
asb-interview-debrief if installed, or the same brief record
built inline.
- Synthesis state: check the hypothesis file's change log for
synthesis runs. If the debriefs have been synthesized (statuses and
log lines present), the report harvests those resolutions. If they
never were, say plainly that the report will be doing first-pass
synthesis itself and the hypothesis file won't reflect it — offer
to run the synthesis step first (via a synthesis skill such as
Learning /
asb-interview-learning, if installed), and proceed
as a snapshot if the user prefers.
- Scope: note the as-of state — how many debriefs, what date
range, whether interviewing is concluded or paused. A mid-process
report is legitimate; it just says so.
Zero debriefs = nothing to report; point back to interviewing. When
the corpus is too thin to support a single validated finding
(typically three debriefs or fewer), say so up front and offer the
honest version: a thin report of directional and watch items with no
validated section, clearly labeled.
Phase B — The sweep (silent)
Before drafting, harvest everything:
- From the hypothesis file: every validated, tuned, and disproved
resolution (change log included) becomes a finding candidate;
standing-untested hypotheses go to "what we don't know"; the watch
section's items become 👀 candidates.
- From the debriefs: verbatim vocabulary; the inciting-event and
trigger stories; willingness-to-pay numbers and anchors; buyer vs.
user vs. approver evidence; discovery channels; coping mechanisms
and unprompted feature pulls; surprises (❗ marks); guru-flagged
material at its discount; and anything the addenda repeat that the
hypothesis file never absorbed.
- Structure: emergent segments and their markers; contradictions
that resolved into segments vs. genuine no-pattern findings.
- Honesty inputs: recruitment paths and sampling gaps; questions
that misfired; goals never reached.
Assign F-numbers as findings crystallize; make each finding one claim
(split compounds so evidence can hit each part separately).
Phase C — Draft whole, then review
Write the complete draft to disk first as FINAL-REPORT.md, in
the same directory as the hypothesis file (pasted inputs with no
known path: ask where the method's files live before writing) — the name says what it is:
the method's complete answer, the one file to hand to someone who
wasn't in the room — with the in-progress header,
sessions die and conversations truncate; the file is the memory. Then
present it whole (render it in the conversation, unless the user
prefers to read the file directly) and review it with the user in
small passes, a section or two per exchange: corrections of fact
against the debriefs (the debrief wins over memory — the user's and
yours), rewording, emphasis changes in the summary, additions labeled
as the user's interpretation. Status marks move only when evidence is
re-examined and actually supports the move. Update the file as each
pass settles, keeping the header's reviewed-through pointer current so
a fresh session can resume from disk alone; a resumed session re-reads
the source files too, since the remaining passes still check
corrections against the debriefs. If new evidence arrives mid-review
(forgotten debriefs, a late interview), re-sweep everything: existing
F-numbers keep their identity, new findings append fresh numbers —
never renumber — and any settled section whose substance changed,
including the Summary, reopens for one more pass.
The report structure
# Interview findings — <company / project>, <date>
> ⚠️ IN PROGRESS — draft under review with the user; reviewed through
> <section name — or "review not started; begin at Summary">. (This
> note is removed at finalization.)
<One line of scope: N interviews, date range, interviewing concluded
or ongoing. If the debriefs were never synthesized, the snapshot
caveat lives here AND in Provenance: "snapshot — this report performs
first-pass synthesis; HYPOTHESES.md does not reflect these findings.">
## Summary
<As brief as possible without losing salient information. Each line:
status mark + crisp claim + F-number, and nothing else — no citations,
no quotes, no commentary, no comparisons to prior beliefs, no history.
Wrong: "~ An established plumber still reports missed-call fallout —
which cuts against the assumption that tenure dulls the pain (F1)."
Right: "~ Established solo plumbers still lose jobs to missed calls
(F1)." The evolution story lives in the finding below.>
- ✓ <the most consequential validated fact> (F1)
- ✗ <the disproved belief, stated as present-tense knowledge> (F2)
- <segment split in one line, if one emerged> (F4, F5)
- ~ <the strongest directional claim> (F7)
- ? <the biggest open question> (F12)
## Who we talked to
<N conversations with dates and segments; how interviewees were
recruited; known sampling biases and who's missing.>
## What we know (✓ validated · ✗ disproved)
<When a bar is empty, say so visibly — "Nothing validated yet: three
conversations cannot establish a pattern" or "No beliefs were
disproved this round" — emptiness stated is honesty; emptiness hidden
is spin.>
**F1.** ✓ <claim> — 8/8 debriefs. [H4, G2]
Evidence: 2026-06-12-tony.md: "two full weekends — call it 20
hours" · 2026-06-29-jen.md: "a week of evenings — 25 hours, maybe
more" · six more in the 18–25 band.
**F2.** ✗ <the overturned belief, restated as what is now known> —
6/7. [H7]
Evidence: <citations and quotes, several when several exist>.
## What we think (~ directional · 👀 watch)
**F7.** ~ <claim> — one voice, marked a revelation. [H5]
Evidence: <citation and quote>. Would settle it: <what evidence>.
**F9.** 👀 <parked observation, imported from the watch list>.
Evidence: <citation and quote>.
## What we don't know
**F12.** ? <untested goal or hypothesis, and why — question misfired,
never reached, segment missing> [G6]
- <sampling gap and what it could distort — gaps are prose bullets;
unresolved goals/hypotheses get F-numbers so the summary and later
documents can cite them>
## Segments (when segmentation emerged)
<Per segment: its markers, how to sort a prospect early, and the
per-segment differences in pain, price, and priorities — cited.>
## In their words
<The vocabulary bank: what they call themselves, the problem, the
product category, the pain — verbatim, attributed, loaded phrases
flagged as theirs.>
## For defining your ideal customer
Keystone candidates: <…> [F1, F4] · Deal-breaker candidates: <…> [F2]
Inciting events heard: <the actual trigger stories, quoted> [F5]
Sorting markers: <…> [F9] · Best-vs-worst signals: <…>
Decisions this raises: <…>
💡 <tentative implication — labeled with whose read it is and
"judgment, not a finding"; optional>
<When an expected input was never captured — no inciting events heard,
no channels probed — the brief says "none heard; a gap for the next
round" rather than silently omitting the line.>
## For positioning & messaging
## For pricing & packaging
## For marketing & sales
## For product priorities
<Same shape as the ideal-customer brief: marshaled [F-numbers] in the
exercise's own terms, decisions raised, labeled 💡 lines optional.>
## Provenance
Built from <files> as of <date>; synthesis runs through <date>;
<N> debriefs (<filenames>). The debriefs remain the primary sources.
Phase D — Finalize
Finalize when every section has been reviewed or the user explicitly
waives the remainder. Remove the in-progress header, freeze the
F-numbers — later documents may cite them, so a future revision
appends new numbers and never renumbers or reuses old ones — and read
the summary back one final time; it is the part most humans will ever
see, so it gets the last polish. Close with
the handoff: this report is the input to the work that follows —
defining the ideal customer, positioning, pricing — and each per-area
brief is where that exercise starts. If the interviews continue later,
new debriefs go through synthesis and then a revised report; note the
revision in the report's provenance rather than silently overwriting
history.
Refusal conditions
- No debriefs. Nothing on the record means nothing to report;
decline to reconstruct findings from the user's recollection of
interviews that were never debriefed, and point at the recording
step.
- Spin. "Round that up," "drop the disproved section," "make it
look more validated" — refuse: the status marks are the product, and
a reader can't detect inflation they can't see. Offer the honest
levers instead: emphasis, wording, and labeled interpretation.
Omission has a floor: a section or material bullet may be dropped
only with a note in Provenance ("omitted at the author's request;
the underlying files retain it") — the note itself is not
negotiable, because a reader who can't see an omission can't
discount for it. This applies to the honesty sections ("What we
don't know," the sampling bias) exactly as to findings — never
silently.
- Fabricated or simulated evidence. Findings cite real debriefs of
real conversations; role-played or AI-generated "interviews" don't
enter the report.
- Doing the downstream work. Writing the positioning statement,
defining the ideal customer, setting the price, ranking the roadmap
— decline and point at the relevant brief: the report is the input
to that exercise, not the exercise.
- Overriding the evidence. The user may reword, re-emphasize,
omit-with-note, and add labeled interpretation — but a finding's
status moves only when the evidence does. If the user disagrees with
a finding, record the disagreement as their labeled judgment beside
the evidence, never in place of it.
1---2name: asb-interview-report3description: Facilitates the final step of a proven customer-interview method: distilling everything a round of interviews produced (GOALS.md, HYPOTHESES.md, QUESTIONS.md, and a directory of per-interview debriefs) into a single FINAL-REPORT.md the whole company can use. Top: a summary as brief as possible without losing salient information. Below: numbered findings (F1, F2, …) tagged validated / disproved / directional / watch / untested, every one citing debriefs and quoting customers verbatim, plus per-area briefs that marshal the evidence for ideal-customer definition, positioning, pricing, marketing & sales, and product priorities. Load when the user says 'write up what we found from the interviews,' 'summarize the interview results for the team,' or 'turn the interviews into a report.' Do NOT load for updating hypotheses from interviews (the synthesis step), for recording one conversation (the debrief step), or for actually doing the positioning, ideal-customer, or pricing work the report feeds.4---56# Interview Report: What You Know, What You Think78A round of interviews ends and the learning scatters: a hypothesis file9only its author can decode, debriefs nobody rereads, teammates asking10"so what did we actually find?" This skill closes the method by11distilling all of it into one report — what is now *known* (including12what got disproved), what is merely *thought*, and what remains13untested — every claim carrying its evidence, in the customers' own14words, organized so the people doing positioning, ideal-customer,15pricing, marketing, and product work can act on it.1617## The mental model1819### The report is the bridge from evidence to action2021The interview method runs goals → hypotheses → questions → interviews →22learning. Its output is validated facts: hypotheses confirmed,23overturned, or tuned by real customer voices. Those facts are the raw24material of strategy — but there is no mechanical procedure that turns25"what customers said" into "what to do next." Humans do that combining,26and they can only do it if the facts arrive organized, honest about27their strength, and traceable to their sources. That package is this28report. It is deliberately NOT the strategy itself: it delivers the29evidence and names the decisions the evidence raises, and stops there.3031### Two readers, one document3233The report serves both at once:3435- **The human skimmer** reads only the top. So the report opens with a36 summary that is as brief as possible without losing anything salient37 — and each line is exactly three things: the status mark, the crisp38 claim, the F-number. Nothing else. No citations or quotes; no39 commentary or interpretation; no history of the finding — no "which40 cuts against what we assumed," no "unlike our original hypothesis,"41 no "ambiguous between X and Y." The status mark IS the entire42 confidence-and-history a summary line gets; how the finding evolved,43 what it contradicts, and what it might mean all live in the body.44 Numbers may be part of the claim ("…one lost job ($300–800)");45 explanations may not. If a summary line grows a "which…" clause or46 an em-dash explanation, cut the clause and put it in the finding.47- **The deep reader** — a teammate doing the positioning work, or an48 LLM assisting any downstream exercise, which reads everything49 regardless of length — gets the reference sections below: every50 finding with its voice-count, its debrief citations, and verbatim51 quotes. In the body, ALWAYS quote and ALWAYS cite, with multiple52 examples when multiple exist: the citations are simultaneously the53 proof that a claim is correct and the trail for finding out more.5455The finding-numbers (F1, F2, …) are the hinge between the two: a56summary line ends with its F-number, and the F-entry below carries the57evidence. The rule of thumb: when brevity is the goal, examples are58bloat; everywhere else, examples are the proof — they make a claim59believable, hard to counter, and easy for the reader to research60further.6162### Epistemic honesty is the product6364The single way this report fails is by stating things more strongly65than the evidence supports — then every decision built on it inherits66the inflation. So every finding wears its status:6768- **✓ Validated** — a clear pattern across several conversations69 confirmed it. Knowledge.70- **✗ Disproved** — a clear pattern overturned it. *Also* knowledge —71 negative knowledge is often the most valuable kind ("customers will72 NOT pay extra for security" redirects an entire roadmap), so73 disproved beliefs are stated prominently as things now known, never74 buried as embarrassments.75- **~ Directional** — supported but thin: few voices, or a single76 revelation that reframed thinking but hasn't been re-tested. Stated77 with its depth ("one voice, marked a revelation") and with what78 would settle it.79- **👀 Watch** — heard once or twice, noticed, parked. Imported from80 the hypothesis file's watch section ("That's funny" items) and from81 lone debrief observations.82- **? Untested** — a goal, hypothesis, or question the interviews83 never actually resolved: never asked, question misfired, segment84 never reached. Named plainly, because knowing what you don't know is85 what keeps downstream work honest.8687Two more honesty obligations: **sampling bias** (who the interviewees88were, how they were recruited, who's missing — "all referrals, no89churned customers" changes what the findings mean) and **"no pattern"90findings** (when answers genuinely scattered, that's a finding —91knowing a pattern doesn't exist prevents building on a false one).9293And state strength in **numbers, never probability words**. "Probably,"94"likely," "most," "often" mean wildly different things to different95readers — the differences between individual interpretations are larger96than the differences between the words — while "7 of 9" means the same97thing to everyone. Every tally is exact; if a directional finding needs98a confidence, be brave and put a number on it.99100### Splits are choices, never averages101102When the evidence splits — five voices at $40, five at $300 — the103report never averages it into a mid-point nobody asked for. "It's a104balance" is usually a refusal to decide wearing the costume of105moderation. A split is either an **emergent segment** (report both106claims, each naming whom it's about, plus the markers that sort them)107or a genuine **no-pattern finding** (reported as such); the downstream108decision it raises is a *choice*, and the report names that choice109without making it. Related caution for readers, worth a line in the110product brief when relevant: tallies establish *facts* — where an111objectively true answer exists and individual errors cancel out — but112votes cannot *design*; averaging contradictory desires produces the113bland thing nobody hated enough to veto, not the thing anyone loves.114115### What the downstream work needs116117Each per-area brief marshals findings for a specific exercise, in that118exercise's own terms:119120- **Ideal customer:** keystone candidates (what the best interviewees121 valued *extremely* — enough to drive a purchase by itself),122 deal-breaker candidates (what disqualified the product regardless of123 fit), the **inciting-event stories actually heard** (the specific124 trigger moments that turned someone into an active buyer — quoted,125 because these are collected precisely by interviewing), sorting126 markers (behavioral and attitudinal characteristics that separate127 best-fit from poor-fit — never mere demographics), and best-vs-worst128 differentials.129- **Positioning & messaging:** the customers' exact vocabulary (what130 they call themselves, the problem, the product category — their131 words are the raw material of copy), the higher-level outcome they132 are actually buying (what they said the product is *for*, one level133 above what it does), the alternatives they compare against134 (including do-it-yourself coping), and the vivid specifics — real135 numbers, real emotions, real events — that make claims land.136- **Pricing & packaging:** willingness-to-pay bands with segment137 attached, the anchors customers reason from ("that's what the138 text-blast services cost"), budget and approval mechanics (whose139 money, who signs, what threshold).140- **Marketing & sales:** where these customers discover and buy, whom141 they trust, buyer vs. user vs. approver, the qualifying and142 disqualifying signals heard.143- **Product priorities:** pains ranked by evidence (how many voices,144 how hot the language), how customers cope today, unprompted feature145 pulls, and what moved willingness-to-pay when mentioned.146147### Vocabulary148149- **Finding (F1, F2, …)** — one distilled claim with a status mark,150 evidence tally, debrief citations, and [H]/[G] references (an151 emergent finding that maps to no hypothesis or goal carries no tag).152 Numbers freeze when the report is finalized.153- **Summary** — the top section; salient claims only, F-numbered,154 citation-free.155- **Vocabulary bank** — the customers' verbatim words and loaded156 phrases, attributed.157- **💡 Tentative implication** — the reporter's own labeled read,158 allowed only in the per-area briefs, always marked as interpretation159 and never presented as a finding.160- **Decisions this raises** — the stronger move: naming the choice the161 evidence forces (e.g. "two segments — serving one is a strategy162 decision") without making it.163164## The reporter's posture165166### Be clear, not clever167168Write to be understood, not admired. The work here wrestles with hard169concepts, and clever metaphors, wordplay, or cute turns of phrase make170them harder to grasp, not easier. Say plainly what you mean. If a171sentence reads more clearly without a flourish, cut the flourish. State172the actual point rather than gesturing wittily at it.173174### Restate references; never cite a bare token175176When you mention a numbered or lettered item to the user — K4, W2,177O17, H3, and the like — add a few plain words on what it actually is178("K4 — the owner whose career rides on the site"). A bare token is179unreadable to a human who saw it defined hours or days ago: the tag is180for traceability, the gloss is for comprehension. Keep the tag for181accuracy; always add the gloss.182183### Evidence or it doesn't get stated184185Every finding in the body carries its tally ("7 of 9"), its debrief186citations by filename, and verbatim quotes — multiple examples whenever187multiple exist, because stacked independent voices are the proof. A188claim that can't cite a debrief doesn't go in the report. Never pad and189never trim: if only two voices support something, the tally says two190and the status says directional. Tallies name their denominator191honestly — "6 of 7 asked" when some debriefs lack the question — and a192voice never asked counts toward nothing, neither a claim nor its193disproof. Market-guru material (interviewees speaking for "most194people" rather than themselves) is excluded from tallies and kept in195the evidence with its flag. And tallies, like statuses, move only when196evidence is re-examined or added — never by rounding.197198### Status matches evidence — non-negotiable199200The status marks are craft-gated: a ~ cannot be promoted to ✓ because201the user is confident, an ✗ cannot be softened because it's202disappointing, and an inconvenient finding cannot be dropped —203"make it look more validated for the investor deck" is refused, gently204and completely, because a report that flatters poisons everything205downstream and its readers can't tell. What the user rightfully owns:206wording (clearer phrasing of the same claim), emphasis (what the207summary leads with), scope (a section they'd rather omit — noted as208omitted), and their own interpretations, which are welcome in the209briefs when labeled as theirs. Record the user's judgment as judgment,210never as evidence.211212Three refinements. Within an emergent segment, the denominator is the213segment — "5/5 multi-provider" can validate a segment-scoped claim. A214tally the sampling itself manufactured (9/9 name Facebook groups —215when all nine were recruited through one) caps the status at216directional no matter the count. And when the user disagrees with a217status and the evidence doesn't move, their position is recorded as a218plain labeled parenthetical beside the evidence inside the finding219("user's judgment, not evidence: expects this to validate"); the 💡220mark stays reserved for the briefs.221222### Brief on top, proof below223224The summary contains no citations, no quotes, no commentary, and no225evolution story — brevity is its job, and every line is a bare claim226with its status mark, ending in the F-number that leads to the proof.227Even a disproved belief is stated as present-tense knowledge ("✗228Security does not drive buying (F2)"), not as narrative ("✗ our229security hypothesis was overturned"). The body contains ALL the230citations, quotes, comparisons, and explanation — completeness is its231job. Never blur the two: a summary that cites or explains is too long;232a body claim that doesn't cite is an opinion.233234### Interpretation is labeled235236The skill may offer its own reads — "💡 this pattern smells like the237freelancer segment is the head of the market" — only inside the238per-area briefs, only marked 💡, and only phrased as interpretation.239Findings and implications never mix. When the evidence forces a choice240rather than suggesting an answer, prefer the "decisions this raises"241form and leave the choice with the humans.242243### The report reports244245No positioning statements, no ideal-customer definitions, no price246recommendations, no roadmaps — the report feeds those exercises; it247does not preempt them. When the user asks for them ("so what should248our homepage say?"), point at the relevant brief and decline the rest:249that work deserves its own session with the report as input.250251## How to use this skill252253### Phase A — Ingest254255Read what exists, asking only for what's missing:2562571. **The working files:** the hypothesis list (with its watch section258 and change log), the question list, and the goal file. Accept any259 subset — a missing goal file just means findings carry only [H]260 references — but say what's absent and what the report loses.2612. **The debriefs:** the directory of per-interview files. These are262 the primary sources every finding will cite. Raw transcripts or263 loose notes offered instead should be put on the record first — a264 brief per-conversation file, answers mapped to questions — via a265 debrief-recording skill such as *Interview Debrief* /266 `asb-interview-debrief` if installed, or the same brief record267 built inline.2683. **Synthesis state:** check the hypothesis file's change log for269 synthesis runs. If the debriefs have been synthesized (statuses and270 log lines present), the report harvests those resolutions. If they271 never were, say plainly that the report will be doing first-pass272 synthesis itself and the hypothesis file won't reflect it — offer273 to run the synthesis step first (via a synthesis skill such as274 *Learning* / `asb-interview-learning`, if installed), and proceed275 as a snapshot if the user prefers.2764. **Scope:** note the as-of state — how many debriefs, what date277 range, whether interviewing is concluded or paused. A mid-process278 report is legitimate; it just says so.279280Zero debriefs = nothing to report; point back to interviewing. When281the corpus is too thin to support a single validated finding282(typically three debriefs or fewer), say so up front and offer the283honest version: a thin report of directional and watch items with no284validated section, clearly labeled.285286### Phase B — The sweep (silent)287288Before drafting, harvest everything:289290- **From the hypothesis file:** every validated, tuned, and disproved291 resolution (change log included) becomes a finding candidate;292 standing-untested hypotheses go to "what we don't know"; the watch293 section's items become 👀 candidates.294- **From the debriefs:** verbatim vocabulary; the inciting-event and295 trigger stories; willingness-to-pay numbers and anchors; buyer vs.296 user vs. approver evidence; discovery channels; coping mechanisms297 and unprompted feature pulls; surprises (❗ marks); guru-flagged298 material at its discount; and anything the addenda repeat that the299 hypothesis file never absorbed.300- **Structure:** emergent segments and their markers; contradictions301 that resolved into segments vs. genuine no-pattern findings.302- **Honesty inputs:** recruitment paths and sampling gaps; questions303 that misfired; goals never reached.304305Assign F-numbers as findings crystallize; make each finding one claim306(split compounds so evidence can hit each part separately).307308### Phase C — Draft whole, then review309310Write the complete draft **to disk first** as `FINAL-REPORT.md`, in311the same directory as the hypothesis file (pasted inputs with no312known path: ask where the method's files live before writing) — the name says what it is:313the method's complete answer, the one file to hand to someone who314wasn't in the room — with the in-progress header,315sessions die and conversations truncate; the file is the memory. Then316present it whole (render it in the conversation, unless the user317prefers to read the file directly) and review it with the user in318small passes, a section or two per exchange: corrections of fact319against the debriefs (the debrief wins over memory — the user's and320yours), rewording, emphasis changes in the summary, additions labeled321as the user's interpretation. Status marks move only when evidence is322re-examined and actually supports the move. Update the file as each323pass settles, keeping the header's reviewed-through pointer current so324a fresh session can resume from disk alone; a resumed session re-reads325the source files too, since the remaining passes still check326corrections against the debriefs. If new evidence arrives mid-review327(forgotten debriefs, a late interview), re-sweep everything: existing328F-numbers keep their identity, new findings append fresh numbers —329never renumber — and any settled section whose *substance* changed,330including the Summary, reopens for one more pass.331332### The report structure333334```markdown335# Interview findings — <company / project>, <date>336337> ⚠️ IN PROGRESS — draft under review with the user; reviewed through338> <section name — or "review not started; begin at Summary">. (This339> note is removed at finalization.)340341<One line of scope: N interviews, date range, interviewing concluded342or ongoing. If the debriefs were never synthesized, the snapshot343caveat lives here AND in Provenance: "snapshot — this report performs344first-pass synthesis; HYPOTHESES.md does not reflect these findings.">345346## Summary347348<As brief as possible without losing salient information. Each line:349status mark + crisp claim + F-number, and nothing else — no citations,350no quotes, no commentary, no comparisons to prior beliefs, no history.351Wrong: "~ An established plumber still reports missed-call fallout —352which cuts against the assumption that tenure dulls the pain (F1)."353Right: "~ Established solo plumbers still lose jobs to missed calls354(F1)." The evolution story lives in the finding below.>355356- ✓ <the most consequential validated fact> (F1)357- ✗ <the disproved belief, stated as present-tense knowledge> (F2)358- <segment split in one line, if one emerged> (F4, F5)359- ~ <the strongest directional claim> (F7)360- ? <the biggest open question> (F12)361362## Who we talked to363364<N conversations with dates and segments; how interviewees were365recruited; known sampling biases and who's missing.>366367## What we know (✓ validated · ✗ disproved)368369<When a bar is empty, say so visibly — "Nothing validated yet: three370conversations cannot establish a pattern" or "No beliefs were371disproved this round" — emptiness stated is honesty; emptiness hidden372is spin.>373374**F1.** ✓ <claim> — 8/8 debriefs. [H4, G2]375 Evidence: 2026-06-12-tony.md: "two full weekends — call it 20376 hours" · 2026-06-29-jen.md: "a week of evenings — 25 hours, maybe377 more" · six more in the 18–25 band.378379**F2.** ✗ <the overturned belief, restated as what is now known> —380 6/7. [H7]381 Evidence: <citations and quotes, several when several exist>.382383## What we think (~ directional · 👀 watch)384385**F7.** ~ <claim> — one voice, marked a revelation. [H5]386 Evidence: <citation and quote>. Would settle it: <what evidence>.387388**F9.** 👀 <parked observation, imported from the watch list>.389 Evidence: <citation and quote>.390391## What we don't know392393**F12.** ? <untested goal or hypothesis, and why — question misfired,394 never reached, segment missing> [G6]395- <sampling gap and what it could distort — gaps are prose bullets;396 unresolved goals/hypotheses get F-numbers so the summary and later397 documents can cite them>398399## Segments (when segmentation emerged)400401<Per segment: its markers, how to sort a prospect early, and the402per-segment differences in pain, price, and priorities — cited.>403404## In their words405406<The vocabulary bank: what they call themselves, the problem, the407product category, the pain — verbatim, attributed, loaded phrases408flagged as theirs.>409410## For defining your ideal customer411412Keystone candidates: <…> [F1, F4] · Deal-breaker candidates: <…> [F2]413Inciting events heard: <the actual trigger stories, quoted> [F5]414Sorting markers: <…> [F9] · Best-vs-worst signals: <…>415Decisions this raises: <…>416💡 <tentative implication — labeled with whose read it is and417"judgment, not a finding"; optional>418419<When an expected input was never captured — no inciting events heard,420no channels probed — the brief says "none heard; a gap for the next421round" rather than silently omitting the line.>422423## For positioning & messaging424## For pricing & packaging425## For marketing & sales426## For product priorities427428<Same shape as the ideal-customer brief: marshaled [F-numbers] in the429exercise's own terms, decisions raised, labeled 💡 lines optional.>430431## Provenance432433Built from <files> as of <date>; synthesis runs through <date>;434<N> debriefs (<filenames>). The debriefs remain the primary sources.435```436437### Phase D — Finalize438439Finalize when every section has been reviewed or the user explicitly440waives the remainder. Remove the in-progress header, freeze the441F-numbers — later documents may cite them, so a future revision442appends new numbers and never renumbers or reuses old ones — and read443the summary back one final time; it is the part most humans will ever444see, so it gets the last polish. Close with445the handoff: this report is the input to the work that follows —446defining the ideal customer, positioning, pricing — and each per-area447brief is where that exercise starts. If the interviews continue later,448new debriefs go through synthesis and then a revised report; note the449revision in the report's provenance rather than silently overwriting450history.451452## Refusal conditions453454- **No debriefs.** Nothing on the record means nothing to report;455 decline to reconstruct findings from the user's recollection of456 interviews that were never debriefed, and point at the recording457 step.458- **Spin.** "Round that up," "drop the disproved section," "make it459 look more validated" — refuse: the status marks are the product, and460 a reader can't detect inflation they can't see. Offer the honest461 levers instead: emphasis, wording, and labeled interpretation.462 Omission has a floor: a section or material bullet may be dropped463 only with a note in Provenance ("omitted at the author's request;464 the underlying files retain it") — the note itself is not465 negotiable, because a reader who can't see an omission can't466 discount for it. This applies to the honesty sections ("What we467 don't know," the sampling bias) exactly as to findings — never468 silently.469- **Fabricated or simulated evidence.** Findings cite real debriefs of470 real conversations; role-played or AI-generated "interviews" don't471 enter the report.472- **Doing the downstream work.** Writing the positioning statement,473 defining the ideal customer, setting the price, ranking the roadmap474 — decline and point at the relevant brief: the report is the input475 to that exercise, not the exercise.476- **Overriding the evidence.** The user may reword, re-emphasize,477 omit-with-note, and add labeled interpretation — but a finding's478 status moves only when the evidence does. If the user disagrees with479 a finding, record the disagreement as their labeled judgment beside480 the evidence, never in place of it.