forger-labs-hq
- 13 skills
- 0 followers
- 1 day ago last updated
- ▌ Researchforge Run · forger-labs-hqExecute an approved experiment plan through the screening → full benchmark funnel, or resume an interrupted run. Use when the user says run the experiments, or a run was interrupted.
- ▌ Researchforge Plan · forger-labs-hqDesign experiment variants for a hypothesis — write patches, import the plan for validation, and get it approved. Use after a baseline exists, when the user wants to plan or implement experiments.
- ▌ Researchforge Ship · forger-labs-hqShip a validated experiment — clean local branch reconstructed from the baseline, engineering report, and optional draft PR. Use when the user wants the winning change as a branch, a report, or a PR.
- ▌ Researchforge Paper · forger-labs-hqBuild the research package — BibTeX citations, related work, evidence matrix, paper outline, reproducibility bundle, and experiment data. Use when the user wants publication materials or a research write-up bundle.
- ▌ Researchforge Start · forger-labs-hqStart or resume a ResearchForge project — explore a research idea or improve a repository with benchmarked experiments. Use when the user wants to begin research, set up ResearchForge, or asks "where was I?" in an existing project.
- ▌ Researchforge Doctor · forger-labs-hqCheck that ResearchForge's dependencies (git, Python, optionally Docker) are available and explain any failures. Use when setup fails, before starting a project, or when the user asks whether their machine is ready.
- ▌ Researchforge Papers · forger-labs-hqSearch arXiv for papers relevant to the project objective and review what was stored. Use when the user wants literature, related work, or asks what papers ResearchForge found.
- ▌ Researchforge Autorun · forger-labs-hqDrive the autonomous research loop round after round — ask the engine which node to expand, plan there, run, read the result, repeat. Use when the user wants to keep improving a repository over many rounds, or wants autorun without an AI API key.
- ▌ Researchforge Results · forger-labs-hqSummarize a run's results — ranking, Pareto trade-offs, constraint violations, and rejected experiments — grounded strictly in recorded measurements. Use when the user asks how the experiments went or which variant won.
- ▌ Researchforge Baseline · forger-labs-hqSet up the experiment contract and run the frozen baseline benchmark. Use when an improve-repository project needs its evaluation defined, the contract approved, or the baseline measured.
- ▌ Researchforge Validate · forger-labs-hqRun repeated validation benchmarks on a run's finalists so a result can honestly be called validated. Use after a run has a promising winner, or when the user asks to confirm/validate a result.
- ▌ Researchforge Landscape · forger-labs-hqSynthesize stored papers into a research landscape — grouped directions with evidence claims — and import it for validation. Use after papers are stored, when the user wants directions, themes, or a map of the literature.
- ▌ Researchforge Hypotheses · forger-labs-hqGenerate testable, evidence-linked hypotheses from the research landscape and import them for validation. Use after the landscape exists, when the user wants hypotheses, experiment ideas, or "what should we try?".