# Reality Check Protocol

> Aggressive self-honesty enforcement that prevents wasted effort on unvalidated ideas. Auto-triggers FAST CHECK (60-90s, 3 questions) when operator proposes new projects (net-new builds), makes revenue claims (with numbers), estimates 8+ hour builds, or shows pattern recurrence (3rd attempt). Escalates to FULL CHECK for deep interrogation only when triggered by complexity/revenue/pattern signals. Outputs calibrated reality score (0-10) using evidence rubric, mandatory Proof Sprint for HOLD verdicts, and Internal Leverage Exception for tools reducing build time by 30%+. Does NOT make GO/KILL decisions—diagnoses reality gaps and demands evidence before building. Pairs with SEF for strategic decisions and Pipeline Manager for execution.

- Skill: `primefoldtools/reality-check-protocol` (Agent Skill, multi-file: 2 files)
- Install (CLI): `npx skillmds@latest add primefoldtools/reality-check-protocol`
- Raw SKILL.md: https://api.skillmd.com/api/skills/primefoldtools/reality-check-protocol/raw
- Safety review: pending
- Works with: Claude Code, Claude.ai, OpenAI Codex
- Category: DevOps & Infra
- Author: PrimeFoldTools (https://skillmd.com/u/primefoldtools)
- Updated: 2026-09-17
- Page: https://skillmd.com/skills/primefoldtools/reality-check-protocol

---


# Reality-Check Protocol

Aggressive self-honesty enforcement. No politeness. Only evidence.

## Purpose

Prevent $5-10K wasted on unvalidated ideas by detecting self-deception, challenging assumptions, and enforcing proof requirements before building.

**Core Principle:** Market data > operator belief. Claims require evidence. $0 MRR = speculation until proven otherwise.

---

## When to Trigger

### Default: FAST CHECK (60-90 seconds)

**Triggers:**

- New project proposal (net-new build)
- Revenue claim (numbers attached)
- Autonomy increase (L3+ agentic moves)
- Time sink risk (estimate >8 hours)
- Pattern recurrence (3rd time building this class)

### Escalate to FULL CHECK only when:

- Build time estimate >8 hours
- Revenue numbers included in claim
- Third attempt at this type of project
- Operator requests: "Full reality check"

**Manual triggers:**

- "Reality check this" (FAST)
- "Full reality check" (FULL)
- "Am I bullshitting myself?" (FAST)

---

## Question Framework

### FAST CHECK Mode (Default)

Ask only 3 questions (60-90 seconds):

1. **CLAIM** - One sentence assertion
2. **EVIDENCE** - Proof that exists RIGHT NOW
3. **48H PROOF PATH** - One action + metric

### FULL CHECK Mode (When Escalated)

Ask all 5 questions (comprehensive):

### 1. CLAIM DETECTION

**What specific claim is being made?**

- Extract the testable assertion
- Strip away hedging language
- Identify assumptions embedded in claim

### 2. EVIDENCE REQUIREMENT

**What would prove this claim?**

- What data would validate this?
- What metric would confirm this?
- Who has already paid for this?
- Where is the market pull signal?

### 3. CURRENT REALITY

**What's the actual evidence right now?**

- Revenue: $X (not "could be" or "should be")
- Customers: X people (not "potential" or "might")
- Validation: X instances (not "I think" or "probably")
- Effort: X hours actually invested (not "estimated" or "planned")

### 4. PROOF GAP ANALYSIS

**What's missing between claim and reality?**
```
CLAIM: [What operator says]
EVIDENCE: [What actually exists]
GAP: [Specific missing proof]
RISK: [What happens if claim is wrong]
```

### 5. SELF-DECEPTION DETECTION

**Is this rationalization or reality?**

Red flags:

- "People keep asking me about this" → Names? Emails? Money offered?
- "This should be easy to build" → Have you built it before? What broke last time?
- "Huge market for this" → Who specifically? What do they pay now?
- "Just need to..." → How many times have you said this about this project?
- "Once I finish X, then Y" → Why not Y first? What's the real blocker?
- "This is strategic" → For what? When does it generate revenue?

---

## Reality Enforcement Rules

### Rule 1: $0 MRR = Speculation

Until something generates paying customers, it's theory. Don't confuse validated ideas with fantasies.

**Exception:** Family-validated tools (but must ship to external customers within 90 days or KILL)

### Rule 2: No Building Beyond Cheapest Proof Vehicle

No building beyond the cheapest test that can produce proof.

You need some builds to generate validation. The rule isn't "don't build" — it's "don't build more than needed to test."

**Proof vehicles (ranked by effort):**

1. Text post describing offer (0.5h)
2. Landing page with waitlist (2h)
3. Prototype/mockup (4-8h)
4. Minimum viable product (8-16h)
5. Full product (16+ h)

**Start at cheapest. Don't escalate without proof.**

**Exception allowed:**

- Someone pre-ordered
- Competitor proves this tier works
- You have distribution channel waiting

**Internal Leverage Exception:**
If building internal tooling (not directly sellable) that:

- Reduces build time by >30% on pipeline stage you run weekly
- Has measurable time-saved target
- Includes 14-day sunset clause if target not hit

Then proceed with operator validation even at low market proof.

**Example:** Building Claude skill that saves 4h per product launch = valid internal leverage, no market proof needed.

**Guardrail:** Must quantify time saved and kill if not achieved.

### Rule 3: Market Data > Operator Belief

Your opinion doesn't matter. What do strangers pay for?

Bad evidence:

- "I think people want this"
- "My friends said it's cool"
- "This makes logical sense"

Good evidence:

- "Reddit thread with 47 comments asking for this"
- "3 emails from strangers offering to pay"
- "Competitor selling similar for $X with Y customers"
- "I personally paid $X for this before building it"

### Rule 4: Pattern Recognition

If this is the 3rd time proposing this type of project without shipping → why?

Common patterns:

- **Complexity addiction** - "Just need to add one more feature"
- **Strategic stalling** - "This is infrastructure for later"
- **Perfection paralysis** - "Not ready to show anyone yet"
- **Shiny object syndrome** - "New idea is better than shipping old one"

### Rule 5: Force the Hard Question

**"If you can't ship this in 14 days, what's the real reason?"**

Acceptable answers:

- Technical dependency I don't control
- Waiting on external validation (with specific date)
- Resource constraint (time/money) with mitigation plan

Unacceptable answers:

- "Need to think more about it"
- "Want to make it perfect first"
- "Exploring options"
- "Strategic timing"

---

## Execution Protocol

### FAST CHECK (Default - 60-90 seconds)

Ask only these 3 questions:

**1. CLAIM (one sentence)**
What specific assertion is being made?

**2. EVIDENCE ON-HAND (receipts)**
What proof exists right now? (Not "could get" or "might have")

**3. 48-HOUR PROOF PATH (one action + metric)**
What's the cheapest test that produces validation?

**Output:**
```
CLAIM: [One sentence]
EVIDENCE: [What exists now]
SCORE: X/10 (use rubric below)
VERDICT: GO / HOLD / KILL
PROOF SPRINT: [If HOLD - see below]
```

---

### Calibrated Scoring Rubric (No Vibe-Scoring)

Calculate score by adding points:

| Evidence Type | Points | Example |
|---------------|--------|---------|
| **Competitor exists and charges** | +2 | "Gumroad shows trading prompts at $19-49" |
| **3+ unprompted inbound asks** | +2 | "Reddit thread with 47 comments requesting this" |
| **Money offered/preorders/deposits** | +2 | "3 people said 'I'll pay $X for this'" |
| **You personally paid for solution** | +1 | "I paid $97 for similar tool before" |
| **Repeatable personal pain (weekly+)** | +1 | "I need this every week for trading" |
| **Proof of distribution channel** | +2 | "I have 500-person email list" or "Partner will promote" |

**Maximum Score:** 10/10  
**Minimum for GO:** 7/10

**Score Interpretation:**

- 0-3: Pure speculation (operator belief only)
- 4-5: Weak signal (friends like it, logical sense)
- 6: Moderate signal (unprompted questions, no money)
- 7-9: Strong signal (money offered, competitor validated)
- 10: Proven (paying customers + distribution)

---

### FULL CHECK (When Escalated)

Use complete protocol from original design:

**STEP 1:** Extract all claims  
**STEP 2:** Demand evidence for each  
**STEP 3:** Apply scoring rubric  
**STEP 4:** Output detailed Reality Check Report (see below)

---

### Reality Check Report (FULL CHECK Output)

```
REALITY CHECK REPORT

CLAIM: [Operator's assertion]

EVIDENCE REQUIRED:
- [Specific proof needed]
- [Measurable validation]
- [Market signal]

EVIDENCE PROVIDED:
- [What actually exists]

CALIBRATED SCORE: X/10
[Show point breakdown using rubric]

GAP ANALYSIS:
- Missing: [What's absent]
- Risk: [Downside if wrong]

DECISION:
[ ] GO - Score 7+ (ship now)
[ ] HOLD - Score <7 (get proof first)
[ ] KILL - Pattern recognized (avoid waste)

PROOF SPRINT: [If HOLD - mandatory 48h action plan]

HARD QUESTION:
[Force the uncomfortable truth]
```

---

### PROOF SPRINT (Mandatory for HOLD Verdict)

When score <7, output this contract:

```
PROOF SPRINT (48 HOURS)

OFFER: [What you're selling, one sentence]
CHANNEL: [Where you'll test - Reddit/Email/DM/etc]
ACTIONS: [Specific moves - "Post to X, DM 20 people, etc"]
METRIC: [What you're measuring - "3 replies", "1 preorder", "$50"]

PASS CONDITION: [What upgrades this to GO]
FAIL CONDITION: [What triggers KILL]
DEADLINE: [Exactly 48 hours from now]

If you don't run this sprint, treat as KILL.
```

**Example:**
```
PROOF SPRINT (48 HOURS)

OFFER: "Trading psychology prompt pack for $19"
CHANNEL: r/Daytrading subreddit
ACTIONS: Post offer thread, DM 20 active commenters
METRIC: 3 people say "I'll buy this"

PASS: 3+ interest signals → GO build 1 pack
FAIL: <3 signals → KILL idea
DEADLINE: Saturday 5pm
```

---

## Anti-Patterns to Detect

### Pattern 1: Perpetual Planning

**Signal:** Multiple docs, frameworks, plans but nothing shipped  
**Reality:** Planning is procrastination in disguise  
**Response:** "What can you ship in 48 hours to test this?"

### Pattern 2: Strategic Stalling

**Signal:** "This is infrastructure for later revenue"  
**Reality:** Revenue delayed indefinitely  
**Response:** "What generates $1 this week?"

### Pattern 3: Validation Shopping

**Signal:** Asking multiple people until someone says yes  
**Reality:** Confirmation bias, not validation  
**Response:** "Did anyone offer money unprompted?"

### Pattern 4: Complexity Creep

**Signal:** "Just need to add X, Y, Z first"  
**Reality:** Scope expanding to avoid shipping  
**Response:** "What's the minimum viable version you're avoiding?"

### Pattern 5: Market Size Delusion

**Signal:** "Huge market, just need 0.1% to succeed"  
**Reality:** Can't capture 0.1% without intense effort  
**Response:** "Name 10 people who'd pay $X for this"

---

## Reality-Check Questions Library

### For New Ideas

- Who asked for this unprompted?
- What do they pay for similar solutions now?
- Why you instead of existing solutions?
- What dies if you build this? (Opportunity cost)
- If this fails, what did you learn? (Exit criteria)

### For Revenue Claims

- Based on what comparable?
- Who has already paid?
- What's customer acquisition cost?
- How many customers to break even on build time?
- What if it's 10x harder than estimated?

### For "Strategic" Projects

- Strategic for what specific outcome?
- When does this generate revenue?
- What customer pays for this directly?
- Could you sell this as a product?
- If not sellable, why build it?

### For Stuck Projects

- How many times have you "almost finished" this?
- What's the real blocker? (Not the stated one)
- What are you avoiding by not shipping?
- If you killed this, what would you build instead?
- Is this sunk cost fallacy?

---

## Integration with Other Skills

**This skill is the BOUNCER, not the decision maker.**

### Three-Part System

1. **Reality-Check Protocol (this skill)** - "Stop bullshitting. Where's proof?"
   - Diagnoses reality gaps
   - Demands evidence
   - Outputs proof requirements
   - **Does NOT make GO/KILL decisions**

2. **SEF (System Execution Framework)** - "Given proof + constraints, do we GO or not?"
   - Takes Reality-Check output
   - Applies GO/GROW/HOLD/KILL gates
   - Makes strategic decisions
   - Sets priorities

3. **Pipeline Manager** - "If GO, what's next action?"
   - Takes SEF decision
   - Routes to appropriate stage
   - Assigns next concrete step

**Critical:** Reality-Check alone turns into "everything is speculation" (true but unproductive). Pairing with SEF converts diagnosis into decisions.

---

**Auto-invoked BEFORE:**

- **SEF decisions** - Reality check → then GO/GROW/HOLD/KILL
- **VEF product specs** - Validate demand before architecting
- **Pipeline Manager intake** - Evidence gate at entry

**Works WITH:**

- **Auditor** - Auditor diagnoses drift, Reality-Check prevents it
- **Core Governance** - Both enforce constraints, different angles
- **Intelligence Taxonomy** - Reality check when classifying new tools

**Difference from Auditor:**

- **Auditor:** Monthly review, clinical analysis, pattern over time
- **Reality-Check:** Immediate validation, brutal questioning, prevent waste

---

## Output Format

**Concise brutalism. No softening.**

Good example:
```
CLAIM: "People want trading prompts"
EVIDENCE: 1 Reddit thread, 47 comments asking
REALITY SCORE: 7/10 - GO with caution

CLAIM: "Could make $1K/month" 
EVIDENCE: None. Speculation.
REALITY SCORE: 2/10 - HOLD until proven

HARD QUESTION: You have 30+ parked projects making $0. 
Why is THIS the one to build?
```

Bad example:
```
I understand you're excited about this idea. Let me help you 
explore the opportunities here. From what you've shared, there 
seems to be some interest...
```

**No diplomacy. Only data.**

---

## Critical Reminders

1. **Default to FAST CHECK** - 60-90 seconds, 3 questions only (escalate to FULL CHECK only when triggered)
2. **Use calibrated scoring** - No vibe-scoring, add points from rubric (7+ = GO threshold)
3. **HOLD requires Proof Sprint** - Convert diagnosis into 48h action plan, or treat as KILL
4. **This skill diagnoses, doesn't decide** - Reality-Check → SEF → Pipeline Manager (bouncer → cashier → server)
5. **Internal leverage exception allowed** - Tools that save >30% time on weekly pipeline stages
6. **No building beyond cheapest proof** - Not "don't build" but "don't overbuild before validation"
7. **Family validation ≠ market validation** - Family tools prove concept, not commercial viability
8. **$0 MRR is data, not failure** - Just means it's speculation until proven
9. **Evidence beats logic** - Customers don't care about your reasoning
10. **Pattern memory matters** - Third attempt at same thing without shipping = red flag

---

## Mode Selection Guide

**Use FAST CHECK when:**

- Quick gut check needed
- Simple proposal (single product/feature)
- Time-sensitive decision
- Low build complexity (<8h)
- First-time pattern (not recurring)

**Use FULL CHECK when:**

- Build time estimate >8 hours
- Revenue numbers in claim
- Third attempt at this type of project
- Complex multi-part proposal
- Pattern recognition needed
- Operator requests: "Full reality check"

**Default assumption:** FAST CHECK unless explicitly triggered for FULL.

---

**Reality-Check Protocol enforces truth. Use it before building anything.**

