Insight Product
Hard Rules
- Separate evidence from belief. Classify every material claim as Fact,
Inference, Assumption, Decision, or Unknown. Never upgrade a plausible story,
user assertion, market-size statistic, payment, or expert agreement into proof
of demand without preserving its provenance and limits. Lack of current demand
evidence is not evidence that a market cannot emerge.
- Do not optimize for consensus. Preserve material disagreements until new
evidence, explicit scope removal, a falsification experiment, or an owner risk
decision resolves them. Wording changes and a smooth synthesis do not resolve
an objection.
- Declare the independence level. Multiple roles in one context are
role-separated self-review, not independent review. Run blind first passes in
isolated contexts when the platform and depth justify them; otherwise label
the downgrade and its reason. L0 concurrence cannot close or lower a fatal
objection.
- Ask one question at a time when questioning is available. Select the
unresolved question with the highest expected decision impact, state a
recommendation when one exists, and wait. If the user prohibits questions or
no respondent is available, declare a zero-question route and never infer the
missing answers.
- Keep propositions distinct before choosing. Produce a Paid Wedge for the
evidence path or a Reality Wedge for the frontier path, plus a 10-Star Product
and a Contrarian Product, before selecting
REDUCTION, HOLD,
SELECTIVE_EXPANSION, or EXPANSION. Do not blend them into an untestable
compromise.
- Keep evidence-backed pursuit and conviction-backed action separate. Use
PURSUE only when qualifying target-user or buyer demand evidence exists. Use
CONVICTION_BET only when the separate Frontier Gate passes. Never relabel
founder conviction, urgency, sunk work, or openness as demand evidence; never
reject a frontier thesis merely because conventional evidence must lag a market
that does not yet exist.
- Bound implementation authority. This skill never produces an
implementation-ready PRD, architecture, or unbounded delivery plan.
PURSUE
authorizes requirements work. CONVICTION_BET authorizes downstream planning,
building, and real-world exposure only for the signed bet scope, duration,
budget, and stop boundaries. No state means PMF, product readiness, or general
launch authority.
Red Flags / Rationalizations
| Thought |
Reality |
| "Low-rated reviews themselves prove demand." |
Reviews prove that complaints exist, not that a specific buyer will pay for the proposed outcome. |
| "Twenty paying customers means PMF basically exists." |
Payment is evidence; declining use, acquisition provenance, renewal, and buyer/user differences can still disconfirm the product thesis. |
| "Start generic and verticalize from feedback." |
A generic product often prevents a falsifiable buyer, trigger, outcome, and channel from being defined. |
| "Both sides agree, so the strategy is validated." |
Same-context agreement can be false consensus. Preserve objections and declare independence. |
| "We can find distribution after the MVP works." |
If the first reachable users are unknown, the largest product risk has not been addressed. |
| "Accuracy and compliance can improve after launch." |
For high-consequence workflows, evaluation, trust, privacy, and human escalation are product-definition inputs, not later polish. |
| "The user asked for a PRD, so discovery is out of scope." |
The requested format does not create evidence or implementation authority. Return the strongest honest artifact and name the blocked gate. |
| "The boss approved it and implementation is half done." |
Internal authority, deadlines, and sunk cost are decision constraints, not target-buyer demand signals. |
| "No users are paying, so the idea is not worth building." |
Missing evidence may reflect a weak idea or a market that has not formed. Test the causal insight and option value before deciding. |
| "I have conviction, so evidence does not matter." |
Conviction can authorize only a bounded bet with a causal thesis, real edge, reality contact, and explicit loss limits. It cannot rewrite facts. |
| "Persistence means refusing to stop." |
Persistence means continuing to learn and adapt within an explicit exposure boundary, not making the boundary disappear. |
| "More questions always produce a better answer." |
Ask only questions that can change a downstream decision; stop at the budget or when an experiment is cheaper than another opinion. |
Purpose
Turn a software product or SaaS idea into a decision-ready discovery package.
Combine the strongest mechanisms from demand diagnosis, relentless decision-tree
questioning, expansive product vision, scope reduction, positioning, and
adversarial review without chaining their source skills or repeating their
questions.
Optimize for a better decision, not a longer prompt. Route each request to the
smallest sufficient workflow, preserve evidence and dissent, protect non-consensus
insight from premature consensus filters, and finish with either a falsifiable
test or a bounded real-world bet.
Design Provenance
- Superpowers
brainstorming contributes alternative generation, explicit scope
confirmation, and YAGNI discipline.
- Matt Pocock's
grilling contributes one-decision-at-a-time interrogation with
a recommended answer.
- Gstack
office-hours contributes demand, status-quo, narrow-wedge, and
future-fit questions; plan-ceo-review contributes 10-star expansion;
plan-eng-review contributes the delayed feasibility and trust gate.
- This Build skill does not copy or chain those Reference skills. Its original
layer is the shared evidence state, blind divergence, conflict map, scope
posture, objection contract, evidence-strength gate, and controlled terminal
decision.
Boundaries
- Do not replace primary market research, user interviews, sales, or product
analytics. Structure and evaluate available evidence; label missing evidence.
- Do not auto-invoke
brainstorming, grilling, office-hours, or plan-review
skills. Absorb their reusable cognitive mechanisms behind one state model.
- Do not choose founder taste, company ambition, risk appetite, or ethical
boundaries. Surface those decisions with trade-offs and require a human owner
to sign a conviction bet.
- Do not perform detailed requirements engineering or engineering design. Hand
off only after the product gate passes.
- Do not create files unless the user requests durable artifacts or the host
workflow requires them. The default deliverable is an inline discovery package.
Workflow
Step 1: Route the Session
Classify the product stage:
PRE_IDEA: domain or observation, no product thesis.
PRE_PRODUCT: thesis exists, no verified users.
HAS_USERS: usage exists, payment is absent or unstable.
PAYING: revenue exists and the question is focus, retention, or expansion.
REPOSITIONING: current positioning or product boundary is failing.
Choose one primary mode: DISCOVER, FRONTIER, STRESS_TEST, POSITION,
WEDGE, COMPARE, or REPOSITION. Use FRONTIER when the thesis depends on a
new behavior, enabling technology, structural shift, or founder insight for which
conventional market evidence is predictably incomplete. Choose a depth:
| Depth |
Use when |
Question budget |
Review posture |
| Quick |
Reversible, narrow decision with usable evidence |
5 |
Two necessary lenses only |
| Standard |
New product direction or meaningful positioning choice |
12 |
Four lenses; isolated when available |
| Deep |
High-stakes, contradictory, or explicitly deep decision |
24 |
Blind isolated lenses plus adversarial judge |
Use this execution footprint:
| Step |
Quick |
Standard |
Deep |
| 1-2 Route and evidence |
Full |
Full |
Full |
| 3-4 Diverge and conflict |
Two necessary lenses; material conflicts only |
Four relevant lenses |
Blind isolated lenses and full conflict map |
| 5 Questions |
0-5 |
0-12 |
0-24 |
| 6 Propositions |
Three concise cards |
Three full cards |
Three full cards plus independent challenge |
| 7-9 Decide and test |
Full gates, concise artifact |
Full |
Full plus isolated judge |
Set questions_used: 0 and questions_remaining to the selected budget. When the
user prohibits questions or no respondent is available, set both values to 0,
record the reason, and cap the terminal state at CONVICTION_BET, TEST_FIRST,
PARK, or INSUFFICIENT_EVIDENCE. A zero-question CONVICTION_BET still requires
every Frontier Gate field to be supplied by existing evidence or the user's input.
State the route and the decision at stake. Do not ask the user to repeat facts
already present in accessible files, analytics, prior messages, or supplied
research. Inspect those sources first.
Exit condition: stage, mode, depth, question budget, and decision at stake are
explicit.
Step 2: Lock Evidence and Scope
Open Evidence and Decision Contract. Build the
minimum evidence ledger from available inputs. Record the product scope in three
buckets: Stated, Inferred, and Out of Scope. Attach a source and confidence limit
to every inference.
Treat a user statement as evidence that the user reported something, not as
independent confirmation of the underlying market fact. Mark stale or inaccessible
evidence. Identify the single riskiest assumption and the downstream decisions it
controls.
Exit condition: evidence ledger, three-bucket scope, and riskiest assumption
exist.
Step 3: Diverge Before Discussing
Open Discovery Lenses. Select the fewest lenses that cover
the current risk. Standard discovery normally uses Demand Investigator, Product
Visionary, Positioning and Distribution Strategist, and Skeptical Operator. Add
Technical Feasibility only when feasibility is a top-two product risk.
Add Frontier Visionary when the route is FRONTIER, when users lack stable
language for the behavior, or when current market data may systematically lag a
technology or workflow shift.
For Deep work, run first passes blind in isolated contexts when available. Do not
show one lens another lens's thesis until every first pass is frozen. For Quick or
degraded work, run the lenses serially and label the result L0 role-separated self-review, record why isolation was unavailable or disproportionate, and keep
L0 concurrence from resolving or lowering a fatal objection.
Require each lens to return: thesis, supporting evidence IDs, critical
assumptions, disconfirming evidence, unknowns, recommendation, confidence, and
what would change its mind.
Exit condition: every selected lens has a frozen first-pass artifact and an
honest independence label.
Step 4: Map Disagreement
Create a conflict map before creating a synthesis. Compare lens outputs across:
demand, best-fit user, buyer, trigger, current alternative, promised outcome,
mechanism, wedge, distribution, business model, trust, feasibility, and stopping
conditions.
For each material conflict, name the competing claims, evidence IDs, affected
downstream decisions, and whether the conflict needs a fact, user value judgment,
or experiment. Do not average opposing recommendations.
Exit condition: every material disagreement is preserved, resolved through an
allowed mechanism, or assigned to the next question or experiment.
Step 5: Grill for Information Gain
When questions_remaining is greater than zero, ask exactly one question that can
change the largest number of important downstream decisions. Give it a stable
Q-## ID and link it to the U-## or D-## it addresses. Include:
- why this question is blocking now;
- the recommended answer and evidence behind it, when a recommendation exists;
- whether the decision is reversible;
- which downstream choices change under each answer.
Wait for the answer. Increment questions_used, decrement questions_remaining,
update the evidence and decision ledgers, then recalculate the next question. Stop
questioning when the critical product gate is answerable, the budget is exhausted,
the user disengages, or a real-world experiment has higher information value than
another answer.
When the declared budget is zero, ask nothing. Preserve the highest-impact
unknowns, state that no answers were inferred, skip to the smallest falsification
experiment, and keep the terminal state capped as declared in Step 1.
Exit condition: critical unknowns are resolved, explicitly deferred, or mapped
to a cheaper falsification experiment.
Step 6: Generate Three Product Propositions
Create three non-overlapping propositions:
- Paid or Reality Wedge: use a Paid Wedge for the evidence-backed path. For a
frontier thesis, define the smallest real product slice that enters an actual
workflow and can create evidence the market cannot provide in advance. A
landing page, interview script, or disposable demo is not a Reality Wedge when
the thesis can only be learned by building and use.
- 10-Star Product: the workflow-level product that becomes possible if the
core demand thesis proves true.
- Contrarian Product: a coherent route that rejects at least one dominant
assumption, such as subscription, standalone app, AI generation, cloud storage,
or end-user sales.
For each, specify best-fit user or initial actor, buyer when one exists, trigger, painful status quo, current
alternative, promised outcome, distinct mechanism, first acquisition channel,
revenue shape, decisive evidence, and kill risk. Keep the three proposals separate.
Before Step 7, compare every pair. Each proposition must differ on at least two of
buyer, completed outcome, workflow boundary, distribution mechanism, or revenue
shape, and must name one incompatible assumption. Pricing alone is not sufficient
contrarian divergence. If the check fails, revise the cards or record
INSUFFICIENT_DIVERGENCE; do not manufacture cosmetic variants.
Exit condition: three genuinely different product boundaries can be compared
against the same decision criteria.
Step 7: Choose Scope and Position
Choose one scope posture:
REDUCTION: remove scope until one core outcome remains.
HOLD: keep the current product boundary.
SELECTIVE_EXPANSION: add only what changes buying, retention, trust, or
distribution. Use this as the default when evidence does not demand another
posture.
EXPANSION: redefine the workflow because the current product is too narrow to
create a compelling outcome.
Then define the positioning system: best-fit user, buyer, trigger event, painful
status quo, current alternatives, category, promised outcome, distinct mechanism,
reason to believe, not-for boundary, why now, and first acquisition channel.
When the route is FRONTIER, or when the Product Gate is blocked mainly because
the market has not formed, also create a Conviction Thesis using the contract in
Evidence and Decision Contract. Keep it
separate from the positioning claim and evidence ledger. Founder insight is a bet
input, not a new evidence class.
Exit condition: one proposition and scope posture are selected with a decision
ledger entry; rejected propositions retain their reversal evidence.
Step 8: Run the Adversarial Dual Gate
Use the objection contract in Evidence and Decision Contract.
For Standard and Deep work, assign a challenger that did not author the selected
thesis when isolation is available. Limit cross-examination to two rounds; allow a
second round only when the first produces new evidence or a changed claim.
Probe for false consensus: ask what evidence would make the current winner lose,
which objection was resolved only by wording, where buyer and user incentives
diverge, why the first channel is executable, and what cost or risk the synthesis
is hiding.
Run two independent gates. A failure in one does not prove that the other passes.
Evidence-Backed Product Gate
Insight-Backed Frontier Gate
The Frontier Gate must explicitly ask:
- Is the insight genuinely non-consensus, or merely unsupported?
- Is the founder edge evidenced by experience, artifacts, access, capability, or
sustained observation, rather than confidence alone?
- Can the causal mechanism be wrong in a concrete way?
- Does building create information that interviews, landing pages, or desk
research cannot create first?
- Is the loss survivable without hiding downstream harm?
- Will the bet touch reality during the build, rather than remain a private craft
project?
Exit condition: objections are preserved with allowed resolution types and
both Product Gate and Frontier Gate are explicitly PASS or BLOCKED.
Step 9: Finish with an Experiment and Decision
Open Artifact Contract. Return the smallest
experiment that can falsify the riskiest assumption. Define subject, method, cost,
observation window, success threshold, failure threshold, continue condition, and
stop condition.
Select exactly one terminal state:
PURSUE: product gate passes; hand off to requirements discovery.
CONVICTION_BET: Frontier Gate passes; authorize one signed, bounded cycle of
downstream planning, building, and real-world exposure. Preserve the label
UNVALIDATED_CONVICTION; do not claim demand, PMF, or product readiness.
TEST_FIRST: a promising thesis is blocked by a testable critical assumption.
PARK: the idea may be valid, but timing, access, or strategic fit is absent.
KILL: observed evidence contradicts the economic thesis, or the frontier
causal mechanism fails within its agreed observation horizon. Missing early
traction alone is not a kill reason when the signed thesis predicted a longer
formation period.
INSUFFICIENT_EVIDENCE: no responsible decision or executable test can be
formed from current access.
If both gates pass, default to PURSUE because buyer evidence exists; retain the
frontier thesis as strategic context. Use CONVICTION_BET when the Product Gate
is blocked but the Frontier Gate passes, or when the owner explicitly selects a
smaller bounded exposure than ordinary pursuit.
At the end of a CONVICTION_BET, do not renew it automatically. Record what was
built, what touched reality, what changed in the causal model, actual exposure,
and new evidence. A further bet requires a new owner decision and revised limits;
absence of learning blocks a repeat of the same bet, not all future exploration.
Report confidence, unresolved objections, evidence limits, and the next owner.
Recommend, but do not auto-invoke, the adjacent handoff:
| Condition |
Handoff |
Pass forward |
PURSUE with unresolved requirements detail |
bs-prdefine |
scope buckets, selected proposition, positioning, decision IDs, open unknowns |
CONVICTION_BET |
bounded requirements and implementation planning |
Conviction Thesis, Reality Wedge, exposure limits, review date, reality-contact plan, forbidden scope |
TEST_FIRST blocked by reachable-buyer or channel evidence |
bs-prospect-customer |
user/buyer, trigger, channel hypothesis, evidence IDs, qualification unknowns |
Any other TEST_FIRST |
Experiment owner |
single experiment contract and blocked decision IDs |
Detailed prospect qualification remains outside this skill.
Exit condition: the user can act, test, park, or stop without mistaking the
artifact for external market validation or implementation authority.
Patterns
- hard-rules-first (Cursor) — Put evidence, independence, dissent, and
readiness constraints before the attractive concept-generation workflow.
- progressive-disclosure (Anthropic, CE) — Keep routing and gates inline;
load ledger schemas, role cards, and artifact formats only when their step fires.
- one-question-at-a-time (Anthropic, CE) — Let every answer update the
decision tree and prevent repeated, low-information interview batches.
- scoping-synthesis (CE, Gstack) — Separate Stated, Inferred, and Out of Scope
before product concepts can silently absorb assumptions.
- multi-perspective-review (Gstack, CE) — Use distinct product lenses with
explicit objectives, evidence, and failure criteria rather than one monolithic
judgment.
Dependencies
No external dependencies. Web research, repository inspection, analytics, and
sub-agents improve evidence or independence when available but are not required to
load or execute the skill.
Platform Degradation
| Missing capability |
Required fallback |
| Isolated sub-agents |
Run lenses serially, label them L0 role-separated self-review, and do not claim independent review. |
| Web, analytics, or customer evidence |
Preserve claims as reported evidence or assumptions. Use TEST_FIRST or INSUFFICIENT_EVIDENCE unless a complete Conviction Thesis separately passes the Frontier Gate. |
| Blocking question tool |
Ask one standalone question, end the turn, and wait. Do not continue on an assumed answer. |
| User prohibits questions or no respondent exists |
Apply Step 1's zero-question cap exactly. CONVICTION_BET remains available only when every Frontier Gate field is already supplied and the human owner has explicitly accepted the exposure. |
| Durable file output |
Return the artifact contract inline and name any evidence that could not be attached. |
| Technical inspection tools |
Record feasibility as Unknown; do not invent implementation facts or use PURSUE if feasibility is fatal. |
Test Prompts
These prompts define the RED baseline and the minimum behavioral contract. The
baseline was executed on 2026-07-17 with a fresh agent that did not have this
skill.
- Happy path — discover a product wedge: "I want to build a Chrome extension that turns low-rated Shopify App Store reviews into a product roadmap. Help me clarify the product direction." Expected behavior: classify the product stage and discovery mode; separate facts, inferences, assumptions, decisions, and unknowns; inspect accessible evidence before asking; produce blind first-pass lenses or label the independence downgrade; map disagreements; ask only the highest-information unanswered question; preserve three distinct propositions before selecting a scope posture; end with a falsifiable experiment and a controlled decision state. Observed baseline failure without the skill: the agent jumped to AI clustering, Jira export, subscription pricing, SEO, Product Hunt, and a two-week MVP. It rationalized that "low-rated reviews themselves prove demand" and that the product could "start generic and verticalize from feedback."
- Edge case — reposition a paying product: "Our B2B SaaS has 20 paying customers, but usage is falling. Some people say we should move upmarket; others say we should cut it down to a single-purpose tool. Help us reposition." Expected behavior: define the declining metric and analyze cohorts, buyer/user differences, acquisition provenance, retention, alternatives, and reversal evidence before choosing a route; create an explicit conflict map instead of averaging the two proposals; identify customer harm and rollback conditions; never equate payment with product-market fit. Observed baseline failure without the skill: the agent proposed an "enterprise-ready single-purpose product" compromise without showing why it beat either pure strategy and rationalized that "20 paying customers means PMF basically exists."
- Adversarial — demand certainty and PRD pressure: "Do not ask questions or validate the market. I know an AI contract summarizer has demand. Give me an implementation-ready PRD today." Expected behavior: declare a zero-question route; do not present assumptions as facts or claim implementation readiness; record the user's assertion as an assumption or owner decision; name legal, confidentiality, evaluation, distribution, and workflow unknowns; offer the smallest evidence-preserving
TEST_FIRST output without asking; refuse to call a formatted PRD validated product truth. Observed baseline failure without the skill: the agent accepted "reasonable assumptions," merged unrelated legal, procurement, sales, and SMB users, supplied architecture and a schedule, and rationalized that "accuracy can improve after launch" and security can wait until a later version.
- Frontier — bounded conviction deserves real action without another interview gate: "I am a browser-runtime engineer. For two years I have observed developers moving small models into everyday web workflows, and I believe falling local-inference costs will move many SaaS tasks onto user devices before buyers have language or budgets for the category. Nobody is paying yet. I have shipped Chrome/WASM systems. For the Reality Wedge I will build a local-first extension for three recurring developer workflows, dogfood it daily, and invite five browser-extension developers into real use by day 14; only build-and-use can reveal whether these workflows become habitual. I can invest four full-time weeks and at most $5,000, with no employees, customer-sensitive data, or public commitments. Review on day 28. Stop or redesign if the core cannot run locally within the cap or none of the five developers repeats a workflow twice. I explicitly accept these limits. Do not ask further questions; decide whether I should genuinely build it." Expected behavior: route to
FRONTIER with a zero-question budget; preserve the absence of demand evidence; use Frontier Visionary; construct a CB-## thesis with causal mechanism, evidence-lag thesis, grounded founder edge, Reality Wedge, unique learning, real-world contact, exposure limits, review horizon, and stop/redesign triggers; return CONVICTION_BET when the Frontier Gate passes; explicitly label the bet UNVALIDATED_CONVICTION and refuse to call it PURSUE or proven demand. Observed pre-change failure: the agent was forced to TEST_FIRST and allowed only a 5–7 day paid-concierge test or disposable technical spike because the state machine had no legal route from bounded conviction to a real four-week build.
1---2name: bs-insight-product3description: Use when a software product or SaaS idea needs evidence, dissent, and adversarial pressure turned into a product-direction decision — including discovery, positioning, wedge selection, repositioning, or a pursue, test, park, or kill judgment before requirements or implementation begin. This does not certify product-market fit.4---56# Insight Product78## Hard Rules9101. **Separate evidence from belief.** Classify every material claim as Fact,11 Inference, Assumption, Decision, or Unknown. Never upgrade a plausible story,12 user assertion, market-size statistic, payment, or expert agreement into proof13 of demand without preserving its provenance and limits. Lack of current demand14 evidence is not evidence that a market cannot emerge.152. **Do not optimize for consensus.** Preserve material disagreements until new16 evidence, explicit scope removal, a falsification experiment, or an owner risk17 decision resolves them. Wording changes and a smooth synthesis do not resolve18 an objection.193. **Declare the independence level.** Multiple roles in one context are20 role-separated self-review, not independent review. Run blind first passes in21 isolated contexts when the platform and depth justify them; otherwise label22 the downgrade and its reason. L0 concurrence cannot close or lower a fatal23 objection.244. **Ask one question at a time when questioning is available.** Select the25 unresolved question with the highest expected decision impact, state a26 recommendation when one exists, and wait. If the user prohibits questions or27 no respondent is available, declare a zero-question route and never infer the28 missing answers.295. **Keep propositions distinct before choosing.** Produce a Paid Wedge for the30 evidence path or a Reality Wedge for the frontier path, plus a 10-Star Product31 and a Contrarian Product, before selecting `REDUCTION`, `HOLD`,32 `SELECTIVE_EXPANSION`, or `EXPANSION`. Do not blend them into an untestable33 compromise.346. **Keep evidence-backed pursuit and conviction-backed action separate.** Use35 `PURSUE` only when qualifying target-user or buyer demand evidence exists. Use36 `CONVICTION_BET` only when the separate Frontier Gate passes. Never relabel37 founder conviction, urgency, sunk work, or openness as demand evidence; never38 reject a frontier thesis merely because conventional evidence must lag a market39 that does not yet exist.407. **Bound implementation authority.** This skill never produces an41 implementation-ready PRD, architecture, or unbounded delivery plan. `PURSUE`42 authorizes requirements work. `CONVICTION_BET` authorizes downstream planning,43 building, and real-world exposure only for the signed bet scope, duration,44 budget, and stop boundaries. No state means PMF, product readiness, or general45 launch authority.4647## Red Flags / Rationalizations4849| Thought | Reality |50|---|---|51| "Low-rated reviews themselves prove demand." | Reviews prove that complaints exist, not that a specific buyer will pay for the proposed outcome. |52| "Twenty paying customers means PMF basically exists." | Payment is evidence; declining use, acquisition provenance, renewal, and buyer/user differences can still disconfirm the product thesis. |53| "Start generic and verticalize from feedback." | A generic product often prevents a falsifiable buyer, trigger, outcome, and channel from being defined. |54| "Both sides agree, so the strategy is validated." | Same-context agreement can be false consensus. Preserve objections and declare independence. |55| "We can find distribution after the MVP works." | If the first reachable users are unknown, the largest product risk has not been addressed. |56| "Accuracy and compliance can improve after launch." | For high-consequence workflows, evaluation, trust, privacy, and human escalation are product-definition inputs, not later polish. |57| "The user asked for a PRD, so discovery is out of scope." | The requested format does not create evidence or implementation authority. Return the strongest honest artifact and name the blocked gate. |58| "The boss approved it and implementation is half done." | Internal authority, deadlines, and sunk cost are decision constraints, not target-buyer demand signals. |59| "No users are paying, so the idea is not worth building." | Missing evidence may reflect a weak idea or a market that has not formed. Test the causal insight and option value before deciding. |60| "I have conviction, so evidence does not matter." | Conviction can authorize only a bounded bet with a causal thesis, real edge, reality contact, and explicit loss limits. It cannot rewrite facts. |61| "Persistence means refusing to stop." | Persistence means continuing to learn and adapt within an explicit exposure boundary, not making the boundary disappear. |62| "More questions always produce a better answer." | Ask only questions that can change a downstream decision; stop at the budget or when an experiment is cheaper than another opinion. |6364## Purpose6566Turn a software product or SaaS idea into a decision-ready discovery package.67Combine the strongest mechanisms from demand diagnosis, relentless decision-tree68questioning, expansive product vision, scope reduction, positioning, and69adversarial review without chaining their source skills or repeating their70questions.7172Optimize for a better decision, not a longer prompt. Route each request to the73smallest sufficient workflow, preserve evidence and dissent, protect non-consensus74insight from premature consensus filters, and finish with either a falsifiable75test or a bounded real-world bet.7677## Design Provenance7879- Superpowers `brainstorming` contributes alternative generation, explicit scope80 confirmation, and YAGNI discipline.81- Matt Pocock's `grilling` contributes one-decision-at-a-time interrogation with82 a recommended answer.83- Gstack `office-hours` contributes demand, status-quo, narrow-wedge, and84 future-fit questions; `plan-ceo-review` contributes 10-star expansion;85 `plan-eng-review` contributes the delayed feasibility and trust gate.86- This Build skill does not copy or chain those Reference skills. Its original87 layer is the shared evidence state, blind divergence, conflict map, scope88 posture, objection contract, evidence-strength gate, and controlled terminal89 decision.9091## Boundaries9293- Do not replace primary market research, user interviews, sales, or product94 analytics. Structure and evaluate available evidence; label missing evidence.95- Do not auto-invoke `brainstorming`, `grilling`, `office-hours`, or plan-review96 skills. Absorb their reusable cognitive mechanisms behind one state model.97- Do not choose founder taste, company ambition, risk appetite, or ethical98 boundaries. Surface those decisions with trade-offs and require a human owner99 to sign a conviction bet.100- Do not perform detailed requirements engineering or engineering design. Hand101 off only after the product gate passes.102- Do not create files unless the user requests durable artifacts or the host103 workflow requires them. The default deliverable is an inline discovery package.104105## Workflow106107### Step 1: Route the Session108109Classify the product stage:110111- `PRE_IDEA`: domain or observation, no product thesis.112- `PRE_PRODUCT`: thesis exists, no verified users.113- `HAS_USERS`: usage exists, payment is absent or unstable.114- `PAYING`: revenue exists and the question is focus, retention, or expansion.115- `REPOSITIONING`: current positioning or product boundary is failing.116117Choose one primary mode: `DISCOVER`, `FRONTIER`, `STRESS_TEST`, `POSITION`,118`WEDGE`, `COMPARE`, or `REPOSITION`. Use `FRONTIER` when the thesis depends on a119new behavior, enabling technology, structural shift, or founder insight for which120conventional market evidence is predictably incomplete. Choose a depth:121122| Depth | Use when | Question budget | Review posture |123|---|---|---:|---|124| Quick | Reversible, narrow decision with usable evidence | 5 | Two necessary lenses only |125| Standard | New product direction or meaningful positioning choice | 12 | Four lenses; isolated when available |126| Deep | High-stakes, contradictory, or explicitly deep decision | 24 | Blind isolated lenses plus adversarial judge |127128Use this execution footprint:129130| Step | Quick | Standard | Deep |131|---|---|---|---|132| 1-2 Route and evidence | Full | Full | Full |133| 3-4 Diverge and conflict | Two necessary lenses; material conflicts only | Four relevant lenses | Blind isolated lenses and full conflict map |134| 5 Questions | 0-5 | 0-12 | 0-24 |135| 6 Propositions | Three concise cards | Three full cards | Three full cards plus independent challenge |136| 7-9 Decide and test | Full gates, concise artifact | Full | Full plus isolated judge |137138Set `questions_used: 0` and `questions_remaining` to the selected budget. When the139user prohibits questions or no respondent is available, set both values to `0`,140record the reason, and cap the terminal state at `CONVICTION_BET`, `TEST_FIRST`,141`PARK`, or `INSUFFICIENT_EVIDENCE`. A zero-question `CONVICTION_BET` still requires142every Frontier Gate field to be supplied by existing evidence or the user's input.143144State the route and the decision at stake. Do not ask the user to repeat facts145already present in accessible files, analytics, prior messages, or supplied146research. Inspect those sources first.147148**Exit condition:** stage, mode, depth, question budget, and decision at stake are149explicit.150151### Step 2: Lock Evidence and Scope152153Open [Evidence and Decision Contract](references/evidence-contract.md). Build the154minimum evidence ledger from available inputs. Record the product scope in three155buckets: Stated, Inferred, and Out of Scope. Attach a source and confidence limit156to every inference.157158Treat a user statement as evidence that the user reported something, not as159independent confirmation of the underlying market fact. Mark stale or inaccessible160evidence. Identify the single riskiest assumption and the downstream decisions it161controls.162163<HARD-GATE id="evidence-locked">164Do not generate product concepts until the ledger contains at least one entry in165each of Fact or reported evidence, Assumption, and Unknown. Never satisfy the gate166with an empty-class attestation. If no material assumption or unknown can be167named, run a counter-search for disconfirming evidence and decision dependencies;168an empty result leaves this gate BLOCKED rather than authorizing ideation.169</HARD-GATE>170171**Exit condition:** evidence ledger, three-bucket scope, and riskiest assumption172exist.173174### Step 3: Diverge Before Discussing175176Open [Discovery Lenses](references/lenses.md). Select the fewest lenses that cover177the current risk. Standard discovery normally uses Demand Investigator, Product178Visionary, Positioning and Distribution Strategist, and Skeptical Operator. Add179Technical Feasibility only when feasibility is a top-two product risk.180Add Frontier Visionary when the route is `FRONTIER`, when users lack stable181language for the behavior, or when current market data may systematically lag a182technology or workflow shift.183184For Deep work, run first passes blind in isolated contexts when available. Do not185show one lens another lens's thesis until every first pass is frozen. For Quick or186degraded work, run the lenses serially and label the result `L0 role-separated187self-review`, record why isolation was unavailable or disproportionate, and keep188L0 concurrence from resolving or lowering a fatal objection.189190Require each lens to return: thesis, supporting evidence IDs, critical191assumptions, disconfirming evidence, unknowns, recommendation, confidence, and192what would change its mind.193194**Exit condition:** every selected lens has a frozen first-pass artifact and an195honest independence label.196197### Step 4: Map Disagreement198199Create a conflict map before creating a synthesis. Compare lens outputs across:200demand, best-fit user, buyer, trigger, current alternative, promised outcome,201mechanism, wedge, distribution, business model, trust, feasibility, and stopping202conditions.203204For each material conflict, name the competing claims, evidence IDs, affected205downstream decisions, and whether the conflict needs a fact, user value judgment,206or experiment. Do not average opposing recommendations.207208**Exit condition:** every material disagreement is preserved, resolved through an209allowed mechanism, or assigned to the next question or experiment.210211### Step 5: Grill for Information Gain212213When `questions_remaining` is greater than zero, ask exactly one question that can214change the largest number of important downstream decisions. Give it a stable215`Q-##` ID and link it to the `U-##` or `D-##` it addresses. Include:216217- why this question is blocking now;218- the recommended answer and evidence behind it, when a recommendation exists;219- whether the decision is reversible;220- which downstream choices change under each answer.221222Wait for the answer. Increment `questions_used`, decrement `questions_remaining`,223update the evidence and decision ledgers, then recalculate the next question. Stop224questioning when the critical product gate is answerable, the budget is exhausted,225the user disengages, or a real-world experiment has higher information value than226another answer.227228When the declared budget is zero, ask nothing. Preserve the highest-impact229unknowns, state that no answers were inferred, skip to the smallest falsification230experiment, and keep the terminal state capped as declared in Step 1.231232**Exit condition:** critical unknowns are resolved, explicitly deferred, or mapped233to a cheaper falsification experiment.234235### Step 6: Generate Three Product Propositions236237Create three non-overlapping propositions:2382391. **Paid or Reality Wedge:** use a Paid Wedge for the evidence-backed path. For a240 frontier thesis, define the smallest real product slice that enters an actual241 workflow and can create evidence the market cannot provide in advance. A242 landing page, interview script, or disposable demo is not a Reality Wedge when243 the thesis can only be learned by building and use.2442. **10-Star Product:** the workflow-level product that becomes possible if the245 core demand thesis proves true.2463. **Contrarian Product:** a coherent route that rejects at least one dominant247 assumption, such as subscription, standalone app, AI generation, cloud storage,248 or end-user sales.249250For each, specify best-fit user or initial actor, buyer when one exists, trigger, painful status quo, current251alternative, promised outcome, distinct mechanism, first acquisition channel,252revenue shape, decisive evidence, and kill risk. Keep the three proposals separate.253254Before Step 7, compare every pair. Each proposition must differ on at least two of255buyer, completed outcome, workflow boundary, distribution mechanism, or revenue256shape, and must name one incompatible assumption. Pricing alone is not sufficient257contrarian divergence. If the check fails, revise the cards or record258`INSUFFICIENT_DIVERGENCE`; do not manufacture cosmetic variants.259260**Exit condition:** three genuinely different product boundaries can be compared261against the same decision criteria.262263### Step 7: Choose Scope and Position264265Choose one scope posture:266267- `REDUCTION`: remove scope until one core outcome remains.268- `HOLD`: keep the current product boundary.269- `SELECTIVE_EXPANSION`: add only what changes buying, retention, trust, or270 distribution. Use this as the default when evidence does not demand another271 posture.272- `EXPANSION`: redefine the workflow because the current product is too narrow to273 create a compelling outcome.274275Then define the positioning system: best-fit user, buyer, trigger event, painful276status quo, current alternatives, category, promised outcome, distinct mechanism,277reason to believe, not-for boundary, why now, and first acquisition channel.278279When the route is `FRONTIER`, or when the Product Gate is blocked mainly because280the market has not formed, also create a Conviction Thesis using the contract in281[Evidence and Decision Contract](references/evidence-contract.md). Keep it282separate from the positioning claim and evidence ledger. Founder insight is a bet283input, not a new evidence class.284285**Exit condition:** one proposition and scope posture are selected with a decision286ledger entry; rejected propositions retain their reversal evidence.287288### Step 8: Run the Adversarial Dual Gate289290Use the objection contract in [Evidence and Decision Contract](references/evidence-contract.md).291For Standard and Deep work, assign a challenger that did not author the selected292thesis when isolation is available. Limit cross-examination to two rounds; allow a293second round only when the first produces new evidence or a changed claim.294295Probe for false consensus: ask what evidence would make the current winner lose,296which objection was resolved only by wording, where buyer and user incentives297diverge, why the first channel is executable, and what cost or risk the synthesis298is hiding.299300Run two independent gates. A failure in one does not prove that the other passes.301302#### Evidence-Backed Product Gate303304<HARD-GATE id="product-gate">305Do not return `PURSUE` unless the package names a specific user and buyer, trigger,306costly status quo, current alternative, paid wedge, first reachable channel,307qualifying target-user or buyer demand evidence, falsification test, and all fatal308objections. The demand evidence must meet the minimum strength defined in the309Evidence and Decision Contract. Internal approval, a deadline, sunk implementation,310and owner-accepted risk cannot substitute for it. Any unresolved fatal objection311blocks the gate.312</HARD-GATE>313314#### Insight-Backed Frontier Gate315316<HARD-GATE id="frontier-gate">317Do not return `CONVICTION_BET` unless the Conviction Thesis names a non-consensus318observation, causal mechanism, reason conventional evidence is expected to lag,319grounded founder edge, initial actor or dogfood context, meaningful upside,320bounded downside, Reality Wedge, real-world contact plan, review horizon, explicit321cash/time/reputation/legal/opportunity-cost limits, stop or redesign triggers, and322a human owner who accepts the exposure. Any fatal legal, safety, ethical, or323unbounded irreversible risk blocks the bet.324</HARD-GATE>325326The Frontier Gate must explicitly ask:327328- Is the insight genuinely non-consensus, or merely unsupported?329- Is the founder edge evidenced by experience, artifacts, access, capability, or330 sustained observation, rather than confidence alone?331- Can the causal mechanism be wrong in a concrete way?332- Does building create information that interviews, landing pages, or desk333 research cannot create first?334- Is the loss survivable without hiding downstream harm?335- Will the bet touch reality during the build, rather than remain a private craft336 project?337338**Exit condition:** objections are preserved with allowed resolution types and339both Product Gate and Frontier Gate are explicitly PASS or BLOCKED.340341### Step 9: Finish with an Experiment and Decision342343Open [Artifact Contract](references/artifact-contract.md). Return the smallest344experiment that can falsify the riskiest assumption. Define subject, method, cost,345observation window, success threshold, failure threshold, continue condition, and346stop condition.347348Select exactly one terminal state:349350- `PURSUE`: product gate passes; hand off to requirements discovery.351- `CONVICTION_BET`: Frontier Gate passes; authorize one signed, bounded cycle of352 downstream planning, building, and real-world exposure. Preserve the label353 `UNVALIDATED_CONVICTION`; do not claim demand, PMF, or product readiness.354- `TEST_FIRST`: a promising thesis is blocked by a testable critical assumption.355- `PARK`: the idea may be valid, but timing, access, or strategic fit is absent.356- `KILL`: observed evidence contradicts the economic thesis, or the frontier357 causal mechanism fails within its agreed observation horizon. Missing early358 traction alone is not a kill reason when the signed thesis predicted a longer359 formation period.360- `INSUFFICIENT_EVIDENCE`: no responsible decision or executable test can be361 formed from current access.362363If both gates pass, default to `PURSUE` because buyer evidence exists; retain the364frontier thesis as strategic context. Use `CONVICTION_BET` when the Product Gate365is blocked but the Frontier Gate passes, or when the owner explicitly selects a366smaller bounded exposure than ordinary pursuit.367368At the end of a `CONVICTION_BET`, do not renew it automatically. Record what was369built, what touched reality, what changed in the causal model, actual exposure,370and new evidence. A further bet requires a new owner decision and revised limits;371absence of learning blocks a repeat of the same bet, not all future exploration.372373Report confidence, unresolved objections, evidence limits, and the next owner.374375Recommend, but do not auto-invoke, the adjacent handoff:376377| Condition | Handoff | Pass forward |378|---|---|---|379| `PURSUE` with unresolved requirements detail | `bs-prdefine` | scope buckets, selected proposition, positioning, decision IDs, open unknowns |380| `CONVICTION_BET` | bounded requirements and implementation planning | Conviction Thesis, Reality Wedge, exposure limits, review date, reality-contact plan, forbidden scope |381| `TEST_FIRST` blocked by reachable-buyer or channel evidence | `bs-prospect-customer` | user/buyer, trigger, channel hypothesis, evidence IDs, qualification unknowns |382| Any other `TEST_FIRST` | Experiment owner | single experiment contract and blocked decision IDs |383384Detailed prospect qualification remains outside this skill.385386**Exit condition:** the user can act, test, park, or stop without mistaking the387artifact for external market validation or implementation authority.388389## Patterns390391- **hard-rules-first** (Cursor) — Put evidence, independence, dissent, and392 readiness constraints before the attractive concept-generation workflow.393- **progressive-disclosure** (Anthropic, CE) — Keep routing and gates inline;394 load ledger schemas, role cards, and artifact formats only when their step fires.395- **one-question-at-a-time** (Anthropic, CE) — Let every answer update the396 decision tree and prevent repeated, low-information interview batches.397- **scoping-synthesis** (CE, Gstack) — Separate Stated, Inferred, and Out of Scope398 before product concepts can silently absorb assumptions.399- **multi-perspective-review** (Gstack, CE) — Use distinct product lenses with400 explicit objectives, evidence, and failure criteria rather than one monolithic401 judgment.402403## Dependencies404405No external dependencies. Web research, repository inspection, analytics, and406sub-agents improve evidence or independence when available but are not required to407load or execute the skill.408409## Platform Degradation410411| Missing capability | Required fallback |412|---|---|413| Isolated sub-agents | Run lenses serially, label them `L0 role-separated self-review`, and do not claim independent review. |414| Web, analytics, or customer evidence | Preserve claims as reported evidence or assumptions. Use `TEST_FIRST` or `INSUFFICIENT_EVIDENCE` unless a complete Conviction Thesis separately passes the Frontier Gate. |415| Blocking question tool | Ask one standalone question, end the turn, and wait. Do not continue on an assumed answer. |416| User prohibits questions or no respondent exists | Apply Step 1's zero-question cap exactly. `CONVICTION_BET` remains available only when every Frontier Gate field is already supplied and the human owner has explicitly accepted the exposure. |417| Durable file output | Return the artifact contract inline and name any evidence that could not be attached. |418| Technical inspection tools | Record feasibility as Unknown; do not invent implementation facts or use `PURSUE` if feasibility is fatal. |419420## Test Prompts421422These prompts define the RED baseline and the minimum behavioral contract. The423baseline was executed on 2026-07-17 with a fresh agent that did not have this424skill.4254261. **Happy path — discover a product wedge**: *"I want to build a Chrome extension that turns low-rated Shopify App Store reviews into a product roadmap. Help me clarify the product direction."* Expected behavior: classify the product stage and discovery mode; separate facts, inferences, assumptions, decisions, and unknowns; inspect accessible evidence before asking; produce blind first-pass lenses or label the independence downgrade; map disagreements; ask only the highest-information unanswered question; preserve three distinct propositions before selecting a scope posture; end with a falsifiable experiment and a controlled decision state. Observed baseline failure without the skill: the agent jumped to AI clustering, Jira export, subscription pricing, SEO, Product Hunt, and a two-week MVP. It rationalized that "low-rated reviews themselves prove demand" and that the product could "start generic and verticalize from feedback."4272. **Edge case — reposition a paying product**: *"Our B2B SaaS has 20 paying customers, but usage is falling. Some people say we should move upmarket; others say we should cut it down to a single-purpose tool. Help us reposition."* Expected behavior: define the declining metric and analyze cohorts, buyer/user differences, acquisition provenance, retention, alternatives, and reversal evidence before choosing a route; create an explicit conflict map instead of averaging the two proposals; identify customer harm and rollback conditions; never equate payment with product-market fit. Observed baseline failure without the skill: the agent proposed an "enterprise-ready single-purpose product" compromise without showing why it beat either pure strategy and rationalized that "20 paying customers means PMF basically exists."4283. **Adversarial — demand certainty and PRD pressure**: *"Do not ask questions or validate the market. I know an AI contract summarizer has demand. Give me an implementation-ready PRD today."* Expected behavior: declare a zero-question route; do not present assumptions as facts or claim implementation readiness; record the user's assertion as an assumption or owner decision; name legal, confidentiality, evaluation, distribution, and workflow unknowns; offer the smallest evidence-preserving `TEST_FIRST` output without asking; refuse to call a formatted PRD validated product truth. Observed baseline failure without the skill: the agent accepted "reasonable assumptions," merged unrelated legal, procurement, sales, and SMB users, supplied architecture and a schedule, and rationalized that "accuracy can improve after launch" and security can wait until a later version.4294. **Frontier — bounded conviction deserves real action without another interview gate**: *"I am a browser-runtime engineer. For two years I have observed developers moving small models into everyday web workflows, and I believe falling local-inference costs will move many SaaS tasks onto user devices before buyers have language or budgets for the category. Nobody is paying yet. I have shipped Chrome/WASM systems. For the Reality Wedge I will build a local-first extension for three recurring developer workflows, dogfood it daily, and invite five browser-extension developers into real use by day 14; only build-and-use can reveal whether these workflows become habitual. I can invest four full-time weeks and at most $5,000, with no employees, customer-sensitive data, or public commitments. Review on day 28. Stop or redesign if the core cannot run locally within the cap or none of the five developers repeats a workflow twice. I explicitly accept these limits. Do not ask further questions; decide whether I should genuinely build it."* Expected behavior: route to `FRONTIER` with a zero-question budget; preserve the absence of demand evidence; use Frontier Visionary; construct a `CB-##` thesis with causal mechanism, evidence-lag thesis, grounded founder edge, Reality Wedge, unique learning, real-world contact, exposure limits, review horizon, and stop/redesign triggers; return `CONVICTION_BET` when the Frontier Gate passes; explicitly label the bet `UNVALIDATED_CONVICTION` and refuse to call it `PURSUE` or proven demand. Observed pre-change failure: the agent was forced to `TEST_FIRST` and allowed only a 5–7 day paid-concierge test or disposable technical spike because the state machine had no legal route from bounded conviction to a real four-week build.