Research Proposal Writer
A proposal is not a paper: there are no results to defend, so it is judged entirely on plausibility and feasibility. The register this forces is specific — you promise questions and deliverables, never outcomes. "We will determine whether X holds under Y, and either answer is publishable" is fundable; "we will improve SOTA by 3%" is a red flag, because nobody can promise a result they haven't produced.
Condensed source material (Heilmeier Catechism, SSRC, MIT Comm Lab, Peyton Jones, Ten Simple Rules) lives in references/proposal-sources.md. Consult it when drafting; the durable principles are baked into this file.
When NOT to use this skill (route out first)
- The gap itself is shaky or unmotivated →
gap-motivation-builder first; there is nothing to propose yet.
- The solution direction hasn't survived any attack →
research-idea-stress-test before investing in prose.
- The user wants the evaluation/comparison plan designed in depth →
benchmark-and-baseline-selector.
- The draft is finished and the user wants a verdict as their real reader would give it →
professor-critic.
- The user needs the SOP/personal essay for an application →
apply-sop-writer.
- The proposal is done and needs defense slides →
deck-beamer-proposal-report.
Intake (one batch, then proceed)
Ask once, together:
- Mode — PhD proposal (prospective supervisor/committee/fellowship, multi-year horizon) or next-paper pitch (current advisor, one paper cycle)?
- Reader and constraints — who reads it, page/word limit, any mandated sections (e.g., NSF Intellectual Merit / Broader Impacts; EU Excellence / Impact / Implementation). Treat program-specific limits and criteria wording as volatile: use what the user provides, and mark anything uncertain "verify against the live call" rather than asserting it.
- What exists — a draft to refine, rough notes, or just an idea?
Precondition: if the user cannot state the research question and why current approaches fail at it, stop — route to gap-motivation-builder and say why. A proposal written around an unmotivated gap wastes the polish.
The three reviewer questions
Every reviewer scans the proposal for exactly three answers (Przeworski & Salomon):
- What will we learn that we do not know now?
- Why is it worth knowing?
- How will we know the conclusions are valid?
Reviewers are overloaded and often outside your subfield; they will not dig for hidden answers. The opening paragraph — first page at most — must answer question 1 crisply, with no jargon. A clearly posed non-rhetorical question or a bold central claim both work. Leave the reader one memorable message: you want to be "the one who claims X" in the committee discussion, not "the one from that university."
The shared skeleton
Both modes fill the same skeleton, which maps onto the Heilmeier Catechism:
| Section |
Answers (Heilmeier) |
| 1. Problem & objective |
What are you trying to do, in plain language? Why is it hard? |
| 2. State of the art & gap |
How is it done today; what are the limits of current practice? |
| 3. Proposed approach |
What is new, and why do you think it will succeed? |
| 4. Workplan & milestones |
How long will it take? What are the mid-term and final "exams"? |
| 5. Evaluation plan |
How will success be measured? |
| 6. Risks & fallbacks |
What are the risks — and what still gets learned if the main idea fails? |
| 7. Expected contributions & payoff |
Who cares; what difference does success make? |
| 8. Fit & resources (PhD mode only) |
Why this supervisor/lab/program specifically? |
Methodology (sections 3–5) is an argument, not an agenda: a list of tasks does not prove the tasks add up to the best attack on the problem. Say what you will actually do — the datasets, instruments, proofs, experiments, analysis techniques — and why that sequence answers the question. Most proposals fail because they leave the reviewer wondering what the applicant will actually do.
Mode calibration
PhD mode (multi-year, prospective reader)
- Scope: roughly 3–4 connected publishable units under one thesis question — a thread, not a list of disconnected projects and not one paper stretched over four years.
- Certainty gradient: year 1 concrete (specific first study, named data/methods), later years directional (questions that depend on early findings). A plan equally detailed in year 4 as in year 1 reads as naive.
- It's a demonstration, not a contract: committees fund the person; nobody holds you to the plan. What's being graded is whether you can pose an important question, know the field, and design a credible attack.
- Fit is dependency, not flattery: name what in the supervisor's expertise, equipment, data, or collaborations the plan actually requires. "I admire your work on X" carries no information; "WP2 needs the longitudinal cohort your lab maintains" does.
- Write for a tired, interdisciplinary panel: first page jargon-free; technical depth later or in an appendix.
- Bibliography is a seriousness signal: reviewers use it to check you won't duplicate existing work; one missing directly-relevant reference is costly.
Next-paper mode (one cycle, your own advisor)
- Scope: one falsifiable claim plus the minimal experiment set that would establish it. If it needs three claims, it's two papers.
- Length: 1–2 pages. Your advisor reads it in minutes; density beats completeness.
- Timeline in weeks/months with explicit decision points: "if the pilot doesn't show the effect by week 6, we pivot to the fallback or kill it."
- End with the ask: what you need from the advisor — compute, data access, a go/no-go, co-author time — stated explicitly.
- Skip the fit section and most background; your advisor knows the field. Spend the space on the claim, the evidence plan, and the risks.
Failure-mode checklist
Run this on every draft (refine mode) and on your own output (draft mode):
- Promised outcomes. Any sentence guaranteeing a result ("will outperform", "will achieve") → rewrite as a question or deliverable where either answer is informative.
- "We will explore / investigate / look at X." Not a research operation. Replace with the actual operations: which experiment, which proof strategy, which dataset, which analysis.
- Task list without argument. The workplan must argue why these tasks are the best attack, not just enumerate them.
- Single point of failure. If WP2–4 all die when WP1's idea fails, the plan is a bet, not a plan. Decouple work packages; give each risk a fallback that still yields a contribution.
- Ambition miscalibration — both directions: one paper's worth stretched over years (too thin) or a career's worth in one proposal (not credible). Calibrate to the mode's scope rule above.
- Nothing at stake. The plan tendentiously marches to a preconceived conclusion. There must be a genuine unknown; say what would prove the idea wrong.
- Vague timeline. Milestones must be checkable events ("working prototype evaluated on X by month 9"), and pivots must be named in advance.
- Jargon on page one. The opening must survive a reader from an adjacent field.
- Fit as flattery (PhD mode). See calibration above.
- No evaluation plan. Every proposal carries a minimal + suggested comparison set (baselines/benchmarks or, for theory, the validity test). Route to
benchmark-and-baseline-selector when this needs real design work.
- Buried lede. The three reviewer questions not answerable from the first page.
Process
Draft mode (from notes or an idea)
- Intake batch.
- Produce a skeleton outline: each section from the table above with a 1–2 sentence commitment of what it will say. Flag any section where information is missing with
[NEEDED: ...] — never invent gap claims, resources, or timelines.
- Get the user's approval on the skeleton (this is where scope and ambition get corrected cheaply).
- Write the prose, then run the failure-mode checklist on it and report anything that survives.
Refine mode (existing draft)
- Read the whole draft first.
- Run the failure-mode checklist; rank findings by severity (kills the proposal / weakens it / polish).
- Report the diagnosis before rewriting — never silently rewrite. The user decides what to accept.
- Revise the affected sections, preserving what already works.
Handoffs
Weak gap → gap-motivation-builder. Untested direction → research-idea-stress-test. Evaluation design → benchmark-and-baseline-selector. Finished draft, named-reader verdict → professor-critic. Slides → deck-beamer-proposal-report.
Output format
# [Proposal title]
[Sections per the skeleton, mode-calibrated]
---
## Notes
**Missing information ([NEEDED] items):** ...
**Claims requiring citations or verification:** ...
**Volatile program details to verify against the live call:** ...
**Checklist items that survived (with severity):** ...
Refine mode additionally opens with the ranked diagnosis before any revised text.
Quality bar
After the first page, the reader can state the question, why it matters, and what you will actually do. After the full read, they can name the milestones, the biggest risk and its fallback, and how success will be measured — and they remember one sentence of it the next day.
1---2name: research-proposal-writer3description: Write or refine a forward-looking research proposal — a plan for work not yet done, judged on plausibility and feasibility rather than results. Two modes: a PhD research proposal for a prospective supervisor, admissions committee, or fellowship (what you will do over the next several years), or a next-paper proposal pitched to your current advisor (what the next paper is and how you will get there). Trigger on "refine my research proposal", "research plan for my PhD application", "proposal for my professor about the next paper", "what I'll work on for the next N years", "pitch my next project to my advisor", "is my research plan convincing". NOT the SOP essay (apply-sop-writer), not grading a finished draft against a named reader (professor-critic), not stress-testing the idea itself (research-idea-stress-test), and not building defense slides (deck-beamer-proposal-report).4---56# Research Proposal Writer78A proposal is not a paper: there are no results to defend, so it is judged entirely on plausibility and feasibility. The register this forces is specific — you promise **questions and deliverables**, never outcomes. "We will determine whether X holds under Y, and either answer is publishable" is fundable; "we will improve SOTA by 3%" is a red flag, because nobody can promise a result they haven't produced.910Condensed source material (Heilmeier Catechism, SSRC, MIT Comm Lab, Peyton Jones, Ten Simple Rules) lives in `references/proposal-sources.md`. Consult it when drafting; the durable principles are baked into this file.1112## When NOT to use this skill (route out first)1314- The gap itself is shaky or unmotivated → `gap-motivation-builder` first; there is nothing to propose yet.15- The solution direction hasn't survived any attack → `research-idea-stress-test` before investing in prose.16- The user wants the evaluation/comparison plan designed in depth → `benchmark-and-baseline-selector`.17- The draft is finished and the user wants a verdict as their real reader would give it → `professor-critic`.18- The user needs the SOP/personal essay for an application → `apply-sop-writer`.19- The proposal is done and needs defense slides → `deck-beamer-proposal-report`.2021## Intake (one batch, then proceed)2223Ask once, together:24251. **Mode** — PhD proposal (prospective supervisor/committee/fellowship, multi-year horizon) or next-paper pitch (current advisor, one paper cycle)?262. **Reader and constraints** — who reads it, page/word limit, any mandated sections (e.g., NSF Intellectual Merit / Broader Impacts; EU Excellence / Impact / Implementation). Treat program-specific limits and criteria wording as volatile: use what the user provides, and mark anything uncertain "verify against the live call" rather than asserting it.273. **What exists** — a draft to refine, rough notes, or just an idea?2829**Precondition:** if the user cannot state the research question and why current approaches fail at it, stop — route to `gap-motivation-builder` and say why. A proposal written around an unmotivated gap wastes the polish.3031## The three reviewer questions3233Every reviewer scans the proposal for exactly three answers (Przeworski & Salomon):34351. **What will we learn that we do not know now?**362. **Why is it worth knowing?**373. **How will we know the conclusions are valid?**3839Reviewers are overloaded and often outside your subfield; they will not dig for hidden answers. The opening paragraph — first page at most — must answer question 1 crisply, with no jargon. A clearly posed non-rhetorical question or a bold central claim both work. Leave the reader one memorable message: you want to be "the one who claims X" in the committee discussion, not "the one from that university."4041## The shared skeleton4243Both modes fill the same skeleton, which maps onto the Heilmeier Catechism:4445| Section | Answers (Heilmeier) |46|---|---|47| 1. Problem & objective | What are you trying to do, in plain language? Why is it hard? |48| 2. State of the art & gap | How is it done today; what are the limits of current practice? |49| 3. Proposed approach | What is new, and why do you think it will succeed? |50| 4. Workplan & milestones | How long will it take? What are the mid-term and final "exams"? |51| 5. Evaluation plan | How will success be measured? |52| 6. Risks & fallbacks | What are the risks — and what still gets learned if the main idea fails? |53| 7. Expected contributions & payoff | Who cares; what difference does success make? |54| 8. Fit & resources *(PhD mode only)* | Why this supervisor/lab/program specifically? |5556Methodology (sections 3–5) is **an argument, not an agenda**: a list of tasks does not prove the tasks add up to the best attack on the problem. Say what you will actually *do* — the datasets, instruments, proofs, experiments, analysis techniques — and why that sequence answers the question. Most proposals fail because they leave the reviewer wondering what the applicant will actually do.5758## Mode calibration5960### PhD mode (multi-year, prospective reader)6162- **Scope:** roughly 3–4 connected publishable units under one thesis question — a thread, not a list of disconnected projects and not one paper stretched over four years.63- **Certainty gradient:** year 1 concrete (specific first study, named data/methods), later years directional (questions that depend on early findings). A plan equally detailed in year 4 as in year 1 reads as naive.64- **It's a demonstration, not a contract:** committees fund the *person*; nobody holds you to the plan. What's being graded is whether you can pose an important question, know the field, and design a credible attack.65- **Fit is dependency, not flattery:** name what in the supervisor's expertise, equipment, data, or collaborations the plan actually *requires*. "I admire your work on X" carries no information; "WP2 needs the longitudinal cohort your lab maintains" does.66- **Write for a tired, interdisciplinary panel:** first page jargon-free; technical depth later or in an appendix.67- **Bibliography is a seriousness signal:** reviewers use it to check you won't duplicate existing work; one missing directly-relevant reference is costly.6869### Next-paper mode (one cycle, your own advisor)7071- **Scope:** one falsifiable claim plus the minimal experiment set that would establish it. If it needs three claims, it's two papers.72- **Length:** 1–2 pages. Your advisor reads it in minutes; density beats completeness.73- **Timeline in weeks/months** with explicit decision points: "if the pilot doesn't show the effect by week 6, we pivot to the fallback or kill it."74- **End with the ask:** what you need from the advisor — compute, data access, a go/no-go, co-author time — stated explicitly.75- **Skip** the fit section and most background; your advisor knows the field. Spend the space on the claim, the evidence plan, and the risks.7677## Failure-mode checklist7879Run this on every draft (refine mode) and on your own output (draft mode):80811. **Promised outcomes.** Any sentence guaranteeing a result ("will outperform", "will achieve") → rewrite as a question or deliverable where either answer is informative.822. **"We will explore / investigate / look at X."** Not a research operation. Replace with the actual operations: which experiment, which proof strategy, which dataset, which analysis.833. **Task list without argument.** The workplan must argue *why these tasks are the best attack*, not just enumerate them.844. **Single point of failure.** If WP2–4 all die when WP1's idea fails, the plan is a bet, not a plan. Decouple work packages; give each risk a fallback that still yields a contribution.855. **Ambition miscalibration** — both directions: one paper's worth stretched over years (too thin) or a career's worth in one proposal (not credible). Calibrate to the mode's scope rule above.866. **Nothing at stake.** The plan tendentiously marches to a preconceived conclusion. There must be a genuine unknown; say what would prove the idea wrong.877. **Vague timeline.** Milestones must be checkable events ("working prototype evaluated on X by month 9"), and pivots must be named in advance.888. **Jargon on page one.** The opening must survive a reader from an adjacent field.899. **Fit as flattery** (PhD mode). See calibration above.9010. **No evaluation plan.** Every proposal carries a minimal + suggested comparison set (baselines/benchmarks or, for theory, the validity test). Route to `benchmark-and-baseline-selector` when this needs real design work.9111. **Buried lede.** The three reviewer questions not answerable from the first page.9293## Process9495### Draft mode (from notes or an idea)96971. Intake batch.982. Produce a **skeleton outline**: each section from the table above with a 1–2 sentence commitment of what it will say. Flag any section where information is missing with `[NEEDED: ...]` — never invent gap claims, resources, or timelines.993. Get the user's approval on the skeleton (this is where scope and ambition get corrected cheaply).1004. Write the prose, then run the failure-mode checklist on it and report anything that survives.101102### Refine mode (existing draft)1031041. Read the whole draft first.1052. Run the failure-mode checklist; rank findings by severity (kills the proposal / weakens it / polish).1063. **Report the diagnosis before rewriting** — never silently rewrite. The user decides what to accept.1074. Revise the affected sections, preserving what already works.108109### Handoffs110111Weak gap → `gap-motivation-builder`. Untested direction → `research-idea-stress-test`. Evaluation design → `benchmark-and-baseline-selector`. Finished draft, named-reader verdict → `professor-critic`. Slides → `deck-beamer-proposal-report`.112113## Output format114115```markdown116# [Proposal title]117118[Sections per the skeleton, mode-calibrated]119120---121122## Notes123124**Missing information ([NEEDED] items):** ...125**Claims requiring citations or verification:** ...126**Volatile program details to verify against the live call:** ...127**Checklist items that survived (with severity):** ...128```129130Refine mode additionally opens with the ranked diagnosis before any revised text.131132## Quality bar133134After the first page, the reader can state the question, why it matters, and what you will actually do. After the full read, they can name the milestones, the biggest risk and its fallback, and how success will be measured — and they remember one sentence of it the next day.