Simulate Elite Experts
Applicability Pre-Check (Gate 0)
Before running the full framework, evaluate 3 dimensions to decide whether this tool is appropriate:
Complexity: Does the problem involve multiple stakeholders, competing constraints, or non-obvious tradeoffs?
- Yes (score 1) → proceed.
- No (score 0) → consider skipping; a direct answer may suffice.
Reversibility: Is the decision hard to reverse once made?
- Hard to reverse (score 1) → proceed; structured deliberation is valuable.
- Easy to reverse (score 0) → consider skipping; low-cost decisions rarely need multi-lens analysis.
Information ambiguity: Is there significant uncertainty about facts, outcomes, or stakeholder preferences?
- High ambiguity (score 1) → proceed; the framework surfaces hidden assumptions.
- Low ambiguity (score 0) → consider skipping; the answer may already be clear.
Decision rule:
- Total >= 2 → use the full framework.
- Total = 1 → offer the user a choice between
micro (lightweight) and full framework.
- Total = 0 → recommend skipping; inform the user this question is better answered directly.
If the user explicitly requests the framework regardless of score, proceed but note the pre-check result.
When NOT to Use This Framework
Do not use (or actively recommend against using) this framework for:
- Simple factual queries: "What is the capital of France?" — no perspectives needed.
- Single-correct-answer problems: "Fix this syntax error" — no tradeoff to explore.
- Urgent time-critical decisions: When seconds matter, structured deliberation adds harmful delay.
- Trivial reversible choices: Low-stakes decisions where the cost of being wrong is negligible.
- Pure emotional support: When the user needs empathy, not analysis.
If the user's question falls into these categories, briefly explain why and offer a direct answer instead.
Core Principle
Treat the model as a viewpoint simulator, not as one stable persona.
Use a fixed four-lens dialogue to answer two core questions:
- What would be a good group of people to explore X?
- What would they say?
Fixed Four-Lens Composition (Hard Constraint)
For classic, lean, and deep profiles, always use exactly four roles:
- Real Person A (specific real person)
- Real Person B (specific real person)
- Domain Expert Archetype (abstract role)
- Omniscient Agent Archetype (abstract role)
For micro profile, use exactly two roles:
- Real Person A (specific real person)
- Domain Expert Archetype (abstract role)
Mandatory rules:
- Real Person roles must be concrete, real, named people (not fictional).
- Domain Expert Archetype must be an abstract domain expert role.
- Omniscient Agent Archetype (when present) must be an abstract omniscient intelligence role.
- Role counts and round counts must match the active profile.
- Do not replace this structure with generic "Expert A/B/C" panels.
Real-Person Selection Criteria (Hard Constraint)
For Real Person A and Real Person B, satisfy all criteria:
- Domain relevance: each person must have direct, public work related to the current problem.
- Public-method traceability: each person must have published ideas, frameworks, or decisions that can be inferred.
- Decision-pressure diversity: the two real people must represent different pressures (for example: product speed vs reliability, science vs operations).
- Time relevance: avoid historically famous but currently irrelevant picks unless historical framing is explicitly required.
For each real person, include:
- Selection rationale in one sentence.
- 2-3 public evidence anchors (for example: known books, talks, essays, open-source work, or widely known decision patterns).
Do not pick real people only for fame value.
Do not claim exact quotes unless quoted from a source in the current turn.
Inference Confidence Annotation
For every statement attributed to a real person (Real Person A/B), append an inline confidence tag:
[confidence: high] — the viewpoint closely follows the person's published framework, methods, or repeated public positions.
[confidence: medium] — the viewpoint is a reasonable extrapolation from the person's known work, but not directly stated by them.
[confidence: low] — the viewpoint is speculative; the person has not publicly addressed this specific topic.
Rules:
- Confidence tags are mandatory for Real Person A/B in every dialogue round.
- Confidence tags are not required for abstract roles (Domain Expert Archetype, Omniscient Agent Archetype).
- If a real person's confidence drops to
low in a round, briefly note why (e.g., "topic outside their published scope").
Real-Person Scoring Matrix (Guardrail)
Before finalizing Real Person A/B, score candidates with this matrix.
Per-person dimensions:
- Domain relevance (0-2)
- Public-method traceability (0-2)
- Time relevance (0-2)
Pair dimension:
- Decision-pressure diversity (0-2, pair-level only)
Passing rules:
- Real Person A score >= 5/6.
- Real Person B score >= 5/6.
- Pair diversity score >= 2/2.
- If any rule fails, rerun candidate selection and mark
low-confidence roster if no better pair is available.
Fallback Strategy (When Real-Person Selection Is Unclear)
Use this deterministic fallback order:
- If user names real people, use them unless unsafe or clearly irrelevant.
- If user gives domain but no names, propose three candidate real-person pairs and pick the best pair with rationale.
- If confidence in pair quality is below 0.6, ask user to select one pair before continuing.
- If user does not choose, proceed with the best pair and explicitly mark
low-confidence roster.
Never replace Real Person A/B with fictional characters.
Never collapse to only abstract roles.
Simulation Safety Rules
- For real people, clearly mark outputs as simulated viewpoints inferred from public work.
- Do not claim private access, private intent, or exact quotes.
- Keep analysis decision-oriented, falsifiable, and domain-specific.
Rolling Uncertainty Tracker
Uncertainty tracking is not limited to the final ledger. Apply rolling updates:
- After Round 1: tag each initial position with its evidence basis (
fact, assumption, or speculation).
- After Round 2: record any assumptions that were challenged and whether they survived cross-examination.
- After Round 3: note which revised positions introduced new assumptions or resolved old ones.
- The final Uncertainty Ledger (Section 8) consolidates the rolling tracker into a clean summary.
This ensures uncertainty is visible throughout the dialogue, not hidden until the end.
Output Contract Guardrail (Hard Constraint)
Section and round counts depend on the active profile:
micro: exactly 5 sections, 2 roles, 2 rounds (4 turns total).
lean: exactly 7 sections, 4 roles, 4 rounds (same as classic, but turns compressed to 1-3 sentences).
classic: exactly 7 sections, 4 roles, 4 rounds (16 turns total).
deep: exactly 9 sections, 4 roles, 6 rounds (24 turns total).
Do not add extra top-level sections beyond the profile's contract.
Each dialogue round must contain one turn from each active role.
Preflight checklist (internal; do not output verbatim):
- Role composition matches selected profile.
- Real-person scoring matrix passes.
- Evidence anchors are present for all real people.
- Section count matches profile contract.
- Each round has turns from all active roles.
- Inference confidence tags are present for all real-person turns.
Postflight checklist (internal; do not output verbatim):
- No fabricated direct quotes for real people.
- Moderator synthesis includes recommendation, strongest alternative, preconditions, early warnings, and next actions.
- Uncertainty ledger cleanly separates facts, assumptions, and speculation.
Failure Modes and Recovery Actions
- FM1: Fame-first roster with weak relevance.
- Recovery: rerank candidates using the scoring matrix; replace weakest candidate.
- FM2: Dialogue turns collapse into agreement too early.
- Recovery: enforce at least one direct challenge per role in Round 2.
- FM3: Missing or malformed section structure.
- Recovery: regenerate with strict 7-section scaffold first, then fill content.
- FM4: Actionability gap in synthesis.
- Recovery: add time horizon, trigger indicators, and 1-3 concrete next actions.
- FM5: Speculation leakage.
- Recovery: move uncertain claims to Uncertainty Ledger and add evidence-needed items.
Controlled Execution Profiles (Structure-Preserving)
Profiles adjust depth and structure according to problem complexity.
micro: 2 roles (1 real person + 1 domain expert archetype), 2 dialogue rounds (initial + final), 5 output sections. Use for medium-complexity problems or when pre-check score = 1.
lean: 4 roles, 4 dialogue rounds, 7 sections (same structure as classic). Compress each turn to 1-3 sentences for low-token contexts.
classic (default): 4 roles, 4 dialogue rounds, balanced detail and readability.
deep: 4 roles, 6 dialogue rounds, 9 sections. Adds metrics, counterarguments, failure triggers, stress test, and contingency planning.
Variable round rules:
- Minimum 2 rounds for any profile (initial positions + final statements).
- Rounds 2 (cross-examination) and 3 (revised positions) may be added or removed based on profile.
- Maximum 6 rounds for
deep profile: adds Round 5 (stress test with adversarial scenarios) and Round 6 (contingency planning).
- Each round always includes one turn from every active role.
Profile selection:
- If user specifies a profile, use it.
- If user does not specify, use
classic.
- If applicability pre-check score = 1, suggest
micro or lean.
Micro Profile Output Sections
- Good Group To Explore X (Two-Lens Roster)
- Dialogue Round 1: Initial Positions
- Dialogue Round 2: Final Statements
- Moderator Synthesis
- Uncertainty Ledger
Deep Profile Additional Rounds
- Round 5: Stress Test — each role describes the scenario where their recommendation fails catastrophically.
- Round 6: Contingency Planning — each role proposes a fallback plan triggered by Round 5 failure scenarios.
Required Output Sections (By Profile)
Classic (default) — 7 sections:
- Good Group To Explore X (Four-Lens Roster)
- Dialogue Round 1: Initial Positions
- Dialogue Round 2: Cross-Examination
- Dialogue Round 3: Revised Positions
- Dialogue Round 4: Final Statements
- Moderator Synthesis
- Uncertainty Ledger
Micro — 5 sections:
- Good Group To Explore X (Two-Lens Roster)
- Dialogue Round 1: Initial Positions
- Dialogue Round 2: Final Statements
- Moderator Synthesis
- Uncertainty Ledger
Deep — up to 9 sections:
- Good Group To Explore X (Four-Lens Roster)
- Dialogue Round 1: Initial Positions
- Dialogue Round 2: Cross-Examination
- Dialogue Round 3: Revised Positions
- Dialogue Round 4: Final Statements
- Dialogue Round 5: Stress Test
- Dialogue Round 6: Contingency Planning
- Moderator Synthesis
- Uncertainty Ledger
Do not skip the Roster section or the Uncertainty Ledger in any profile.
Each dialogue round must contain one turn from each active role.
Workflow
- Define decision frame
- Restate question, success criteria, constraints, and time horizon.
- Declare assumptions when context is missing.
- Build four-lens roster
- Select two real people with clear relevance to the problem.
- Explain why each role belongs in the group.
- Score Real Person A/B with the scoring matrix before finalizing.
- Run multi-round dialogue
- Round 1: initial claims.
- Round 2: challenges and tradeoffs.
- Round 3: revised positions after challenge.
- Round 4: final stance and one concrete action.
- Synthesize
- Merge strongest arguments into one recommendation.
- State why it beats the strongest alternative.
- Include preconditions, early warning indicators, and next actions.
- Calibrate uncertainty
- Separate facts, assumptions, and speculation.
- List evidence needed for confidence upgrades.
- Run guardrail self-check
- Validate structure, safety, and actionability before final output.
- User interaction and post-use reflection
- After delivering the output, append the Post-Use Self-Check section.
- If the session is interactive, invite the user to mark which positions they agree/disagree with and why.
User Interaction Guidance
The framework output should not be a passive report. Build in interaction touchpoints:
Pre-dialogue check-in: After presenting the roster (Section 1), pause and ask the user:
- "Do these roles and people look right for your question? Would you swap anyone?"
- If the user confirms, proceed. If the user suggests changes, adjust before running dialogue.
Mid-dialogue checkpoint (optional, for deep profile): After Round 2 (Cross-Examination), briefly ask:
- "Any assumptions you think were missed in the cross-examination?"
Post-output reflection: After the Uncertainty Ledger, always append the Post-Use Self-Check.
These interaction points transform the output from "AI-generated report" to "collaborative thinking scaffold."
Post-Use Self-Check
After the final section (Uncertainty Ledger), always append a self-check block to help the user actively process the output rather than passively accept it.
Template:
### Post-Use Self-Check
1. Before reading this analysis, what was your initial leaning?
2. After reading, has your position changed? If so, which argument was most persuasive?
3. Which assumption in the Uncertainty Ledger concerns you most?
4. What is one piece of evidence you could gather in the next 48 hours to reduce uncertainty?
5. If you had to decide right now, what would you choose and why?
Rules:
- This block is mandatory in all profiles (micro, lean, classic, deep).
- It appears after the Uncertainty Ledger, as a non-numbered appendix (not counted in the main section count).
- Keep it to exactly 5 questions. Do not expand or customize.
Evaluation and Regression
Use:
references/eval-rubric.md for scoring criteria.
references/eval-cases.md for regression test prompts.
references/first-use-guide.md for onboarding new users.
scripts/lint_response.ps1 for hard-gate structure checks on generated outputs.
When updating this skill:
- Run at least 5 cases from
eval-cases.md.
- Ensure every case matches its profile's section and role count requirements.
- Track rubric score before/after edits and avoid regressions (new baseline: 0-20 scale).
- Record outcomes using a compact log: date, cases run, pass rate, avg score, fail reasons.
Output Contract
- Use
references/output-templates.md for English output.
- Use
references/output-templates-zh.md for Chinese output.
- If user asks for brevity, keep all sections required by the active profile and compress each section to 1-3 bullets.
- If using
lean profile, keep all required sections and all active role turns per round.
- If using
micro profile, use the 5-section template with 2 roles.
1---2name: simulate-elite-experts3description: Simulate high-stakes reasoning by modeling how top experts in the relevant domain would think, disagree, and converge on a decision. Use when users ask to role-play strongest minds, compare elite viewpoints, or ask: what would be a good group of people to explore X, and what would they say. Trigger for prompts like "think like world-class experts", "simulate top domain specialists", "role-play strongest domain people", and "use four-lens dialogue".4---5
6# Simulate Elite Experts
7
8## Applicability Pre-Check (Gate 0)
9
10Before running the full framework, evaluate 3 dimensions to decide whether this tool is appropriate:
11
121. **Complexity**: Does the problem involve multiple stakeholders, competing constraints, or non-obvious tradeoffs?
13 - Yes (score 1) → proceed.
14 - No (score 0) → consider skipping; a direct answer may suffice.
15
162. **Reversibility**: Is the decision hard to reverse once made?
17 - Hard to reverse (score 1) → proceed; structured deliberation is valuable.
18 - Easy to reverse (score 0) → consider skipping; low-cost decisions rarely need multi-lens analysis.
19
203. **Information ambiguity**: Is there significant uncertainty about facts, outcomes, or stakeholder preferences?
21 - High ambiguity (score 1) → proceed; the framework surfaces hidden assumptions.
22 - Low ambiguity (score 0) → consider skipping; the answer may already be clear.
23
24**Decision rule**:
25- Total >= 2 → use the full framework.
26- Total = 1 → offer the user a choice between `micro` (lightweight) and full framework.
27- Total = 0 → recommend skipping; inform the user this question is better answered directly.
28
29If the user explicitly requests the framework regardless of score, proceed but note the pre-check result.
30
31## When NOT to Use This Framework
32
33Do not use (or actively recommend against using) this framework for:
34- **Simple factual queries**: "What is the capital of France?" — no perspectives needed.
35- **Single-correct-answer problems**: "Fix this syntax error" — no tradeoff to explore.
36- **Urgent time-critical decisions**: When seconds matter, structured deliberation adds harmful delay.
37- **Trivial reversible choices**: Low-stakes decisions where the cost of being wrong is negligible.
38- **Pure emotional support**: When the user needs empathy, not analysis.
39
40If the user's question falls into these categories, briefly explain why and offer a direct answer instead.
41
42## Core Principle
43
44Treat the model as a viewpoint simulator, not as one stable persona.
45Use a fixed four-lens dialogue to answer two core questions:
461) What would be a good group of people to explore X?
472) What would they say?
48
49## Fixed Four-Lens Composition (Hard Constraint)
50
51For `classic`, `lean`, and `deep` profiles, always use exactly four roles:
521. Real Person A (specific real person)
532. Real Person B (specific real person)
543. Domain Expert Archetype (abstract role)
554. Omniscient Agent Archetype (abstract role)
56
57For `micro` profile, use exactly two roles:
581. Real Person A (specific real person)
592. Domain Expert Archetype (abstract role)
60
61Mandatory rules:
62- Real Person roles must be concrete, real, named people (not fictional).
63- Domain Expert Archetype must be an abstract domain expert role.
64- Omniscient Agent Archetype (when present) must be an abstract omniscient intelligence role.
65- Role counts and round counts must match the active profile.
66- Do not replace this structure with generic "Expert A/B/C" panels.
67
68## Real-Person Selection Criteria (Hard Constraint)
69
70For Real Person A and Real Person B, satisfy all criteria:
71- Domain relevance: each person must have direct, public work related to the current problem.
72- Public-method traceability: each person must have published ideas, frameworks, or decisions that can be inferred.
73- Decision-pressure diversity: the two real people must represent different pressures (for example: product speed vs reliability, science vs operations).
74- Time relevance: avoid historically famous but currently irrelevant picks unless historical framing is explicitly required.
75
76For each real person, include:
77- Selection rationale in one sentence.
78- 2-3 public evidence anchors (for example: known books, talks, essays, open-source work, or widely known decision patterns).
79
80Do not pick real people only for fame value.
81Do not claim exact quotes unless quoted from a source in the current turn.
82
83## Inference Confidence Annotation
84
85For every statement attributed to a real person (Real Person A/B), append an inline confidence tag:
86- `[confidence: high]` — the viewpoint closely follows the person's published framework, methods, or repeated public positions.
87- `[confidence: medium]` — the viewpoint is a reasonable extrapolation from the person's known work, but not directly stated by them.
88- `[confidence: low]` — the viewpoint is speculative; the person has not publicly addressed this specific topic.
89
90Rules:
91- Confidence tags are mandatory for Real Person A/B in every dialogue round.
92- Confidence tags are not required for abstract roles (Domain Expert Archetype, Omniscient Agent Archetype).
93- If a real person's confidence drops to `low` in a round, briefly note why (e.g., "topic outside their published scope").
94
95## Real-Person Scoring Matrix (Guardrail)
96
97Before finalizing Real Person A/B, score candidates with this matrix.
98
99Per-person dimensions:
100- Domain relevance (0-2)
101- Public-method traceability (0-2)
102- Time relevance (0-2)
103
104Pair dimension:
105- Decision-pressure diversity (0-2, pair-level only)
106
107Passing rules:
108- Real Person A score >= 5/6.
109- Real Person B score >= 5/6.
110- Pair diversity score >= 2/2.
111- If any rule fails, rerun candidate selection and mark `low-confidence roster` if no better pair is available.
112
113## Fallback Strategy (When Real-Person Selection Is Unclear)
114
115Use this deterministic fallback order:
1161. If user names real people, use them unless unsafe or clearly irrelevant.
1172. If user gives domain but no names, propose three candidate real-person pairs and pick the best pair with rationale.
1183. If confidence in pair quality is below 0.6, ask user to select one pair before continuing.
1194. If user does not choose, proceed with the best pair and explicitly mark `low-confidence roster`.
120
121Never replace Real Person A/B with fictional characters.
122Never collapse to only abstract roles.
123
124## Simulation Safety Rules
125
126- For real people, clearly mark outputs as simulated viewpoints inferred from public work.
127- Do not claim private access, private intent, or exact quotes.
128- Keep analysis decision-oriented, falsifiable, and domain-specific.
129
130## Rolling Uncertainty Tracker
131
132Uncertainty tracking is not limited to the final ledger. Apply rolling updates:
133- After Round 1: tag each initial position with its evidence basis (`fact`, `assumption`, or `speculation`).
134- After Round 2: record any assumptions that were challenged and whether they survived cross-examination.
135- After Round 3: note which revised positions introduced new assumptions or resolved old ones.
136- The final Uncertainty Ledger (Section 8) consolidates the rolling tracker into a clean summary.
137
138This ensures uncertainty is visible throughout the dialogue, not hidden until the end.
139
140## Output Contract Guardrail (Hard Constraint)
141
142Section and round counts depend on the active profile:
143- `micro`: exactly 5 sections, 2 roles, 2 rounds (4 turns total).
144- `lean`: exactly 7 sections, 4 roles, 4 rounds (same as classic, but turns compressed to 1-3 sentences).
145- `classic`: exactly 7 sections, 4 roles, 4 rounds (16 turns total).
146- `deep`: exactly 9 sections, 4 roles, 6 rounds (24 turns total).
147
148Do not add extra top-level sections beyond the profile's contract.
149Each dialogue round must contain one turn from each active role.
150
151Preflight checklist (internal; do not output verbatim):
1521. Role composition matches selected profile.
1532. Real-person scoring matrix passes.
1543. Evidence anchors are present for all real people.
1554. Section count matches profile contract.
1565. Each round has turns from all active roles.
1576. Inference confidence tags are present for all real-person turns.
158
159Postflight checklist (internal; do not output verbatim):
1601. No fabricated direct quotes for real people.
1612. Moderator synthesis includes recommendation, strongest alternative, preconditions, early warnings, and next actions.
1623. Uncertainty ledger cleanly separates facts, assumptions, and speculation.
163
164## Failure Modes and Recovery Actions
165
166- FM1: Fame-first roster with weak relevance.
167 - Recovery: rerank candidates using the scoring matrix; replace weakest candidate.
168- FM2: Dialogue turns collapse into agreement too early.
169 - Recovery: enforce at least one direct challenge per role in Round 2.
170- FM3: Missing or malformed section structure.
171 - Recovery: regenerate with strict 7-section scaffold first, then fill content.
172- FM4: Actionability gap in synthesis.
173 - Recovery: add time horizon, trigger indicators, and 1-3 concrete next actions.
174- FM5: Speculation leakage.
175 - Recovery: move uncertain claims to Uncertainty Ledger and add evidence-needed items.
176
177## Controlled Execution Profiles (Structure-Preserving)
178
179Profiles adjust depth and structure according to problem complexity.
180
181- `micro`: 2 roles (1 real person + 1 domain expert archetype), 2 dialogue rounds (initial + final), 5 output sections. Use for medium-complexity problems or when pre-check score = 1.
182- `lean`: 4 roles, 4 dialogue rounds, 7 sections (same structure as classic). Compress each turn to 1-3 sentences for low-token contexts.
183- `classic` (default): 4 roles, 4 dialogue rounds, balanced detail and readability.
184- `deep`: 4 roles, 6 dialogue rounds, 9 sections. Adds metrics, counterarguments, failure triggers, stress test, and contingency planning.
185
186Variable round rules:
187- Minimum 2 rounds for any profile (initial positions + final statements).
188- Rounds 2 (cross-examination) and 3 (revised positions) may be added or removed based on profile.
189- Maximum 6 rounds for `deep` profile: adds Round 5 (stress test with adversarial scenarios) and Round 6 (contingency planning).
190- Each round always includes one turn from every active role.
191
192Profile selection:
193- If user specifies a profile, use it.
194- If user does not specify, use `classic`.
195- If applicability pre-check score = 1, suggest `micro` or `lean`.
196
197### Micro Profile Output Sections
198
1991. Good Group To Explore X (Two-Lens Roster)
2002. Dialogue Round 1: Initial Positions
2013. Dialogue Round 2: Final Statements
2024. Moderator Synthesis
2035. Uncertainty Ledger
204
205### Deep Profile Additional Rounds
206
207- Round 5: Stress Test — each role describes the scenario where their recommendation fails catastrophically.
208- Round 6: Contingency Planning — each role proposes a fallback plan triggered by Round 5 failure scenarios.
209
210## Required Output Sections (By Profile)
211
212### Classic (default) — 7 sections:
2131. Good Group To Explore X (Four-Lens Roster)
2142. Dialogue Round 1: Initial Positions
2153. Dialogue Round 2: Cross-Examination
2164. Dialogue Round 3: Revised Positions
2175. Dialogue Round 4: Final Statements
2186. Moderator Synthesis
2197. Uncertainty Ledger
220
221### Micro — 5 sections:
2221. Good Group To Explore X (Two-Lens Roster)
2232. Dialogue Round 1: Initial Positions
2243. Dialogue Round 2: Final Statements
2254. Moderator Synthesis
2265. Uncertainty Ledger
227
228### Deep — up to 9 sections:
2291. Good Group To Explore X (Four-Lens Roster)
2302. Dialogue Round 1: Initial Positions
2313. Dialogue Round 2: Cross-Examination
2324. Dialogue Round 3: Revised Positions
2335. Dialogue Round 4: Final Statements
2346. Dialogue Round 5: Stress Test
2357. Dialogue Round 6: Contingency Planning
2368. Moderator Synthesis
2379. Uncertainty Ledger
238
239Do not skip the Roster section or the Uncertainty Ledger in any profile.
240Each dialogue round must contain one turn from each active role.
241
242## Workflow
243
2441. Define decision frame
245- Restate question, success criteria, constraints, and time horizon.
246- Declare assumptions when context is missing.
247
2482. Build four-lens roster
249- Select two real people with clear relevance to the problem.
250- Explain why each role belongs in the group.
251- Score Real Person A/B with the scoring matrix before finalizing.
252
2533. Run multi-round dialogue
254- Round 1: initial claims.
255- Round 2: challenges and tradeoffs.
256- Round 3: revised positions after challenge.
257- Round 4: final stance and one concrete action.
258
2594. Synthesize
260- Merge strongest arguments into one recommendation.
261- State why it beats the strongest alternative.
262- Include preconditions, early warning indicators, and next actions.
263
2645. Calibrate uncertainty
265- Separate facts, assumptions, and speculation.
266- List evidence needed for confidence upgrades.
267
2686. Run guardrail self-check
269- Validate structure, safety, and actionability before final output.
270
2717. User interaction and post-use reflection
272- After delivering the output, append the Post-Use Self-Check section.
273- If the session is interactive, invite the user to mark which positions they agree/disagree with and why.
274
275## User Interaction Guidance
276
277The framework output should not be a passive report. Build in interaction touchpoints:
278
2791. **Pre-dialogue check-in**: After presenting the roster (Section 1), pause and ask the user:
280 - "Do these roles and people look right for your question? Would you swap anyone?"
281 - If the user confirms, proceed. If the user suggests changes, adjust before running dialogue.
282
2832. **Mid-dialogue checkpoint** (optional, for `deep` profile): After Round 2 (Cross-Examination), briefly ask:
284 - "Any assumptions you think were missed in the cross-examination?"
285
2863. **Post-output reflection**: After the Uncertainty Ledger, always append the Post-Use Self-Check.
287
288These interaction points transform the output from "AI-generated report" to "collaborative thinking scaffold."
289
290## Post-Use Self-Check
291
292After the final section (Uncertainty Ledger), always append a self-check block to help the user actively process the output rather than passively accept it.
293
294Template:
295```
296### Post-Use Self-Check
2971. Before reading this analysis, what was your initial leaning?
2982. After reading, has your position changed? If so, which argument was most persuasive?
2993. Which assumption in the Uncertainty Ledger concerns you most?
3004. What is one piece of evidence you could gather in the next 48 hours to reduce uncertainty?
3015. If you had to decide right now, what would you choose and why?
302```
303
304Rules:
305- This block is mandatory in all profiles (micro, lean, classic, deep).
306- It appears after the Uncertainty Ledger, as a non-numbered appendix (not counted in the main section count).
307- Keep it to exactly 5 questions. Do not expand or customize.
308
309## Evaluation and Regression
310
311Use:
312- `references/eval-rubric.md` for scoring criteria.
313- `references/eval-cases.md` for regression test prompts.
314- `references/first-use-guide.md` for onboarding new users.
315- `scripts/lint_response.ps1` for hard-gate structure checks on generated outputs.
316
317When updating this skill:
318- Run at least 5 cases from `eval-cases.md`.
319- Ensure every case matches its profile's section and role count requirements.
320- Track rubric score before/after edits and avoid regressions (new baseline: 0-20 scale).
321- Record outcomes using a compact log: date, cases run, pass rate, avg score, fail reasons.
322
323## Output Contract
324
325- Use `references/output-templates.md` for English output.
326- Use `references/output-templates-zh.md` for Chinese output.
327- If user asks for brevity, keep all sections required by the active profile and compress each section to 1-3 bullets.
328- If using `lean` profile, keep all required sections and all active role turns per round.
329- If using `micro` profile, use the 5-section template with 2 roles.