Session End Skill
Platform Note: State files (STATE.md, wave-scope.json) live in the platform's native directory:
.claude/(Claude Code),.codex/(Codex CLI), or.cursor/(Cursor IDE). All references to.claude/below should use the platform's state directory. Shared metrics live in.orchestrator/metrics/. Seeskills/_shared/platform-tools.md.
Project-instruction file:
CLAUDE.mdandAGENTS.md(Codex CLI) are transparent aliases — see skills/_shared/instruction-file-resolution.md. All references toCLAUDE.mdin this skill resolve via that precedence rule.
Phase 0: Bootstrap Gate
Read skills/_shared/bootstrap-gate.md and execute the gate check. If the gate is CLOSED, invoke skills/bootstrap/SKILL.md and wait for completion before proceeding. If the gate is OPEN, continue to Phase 1.
Phase 0.5: Parallel-Aware Preamble
Skip silently when
persistence: falsein Session Config.
Before Phase 1, run the parallel-aware preamble per skills/_shared/parallel-aware-preamble.md. The preamble detects other active sessions in the worktree-family via findPeers(repoRoot, { mySessionId }), classifies the caller's mode via classifyMode(callerMode) against the exclusivity-matrix, and fires the appropriate AUQ on conflict.
Outcome handling:
PASS_THROUGH→ continue to Phase 1EXCLUSIVE_BLOCKED→ exit Phase 0 cleanly per the AUQ outcomePROMOTION_OFFER→ user picks Worktree-Promotion (seeparallel-aware-auq.mdoutcome-handling — callsenterWorktree()), in-place + Deviation, or Abbrechen
For session-end specifically: the preamble is DETECTION-ONLY. The lock-release path in later phases keeps its current behavior — releasing the OWN session's lock requires no matrix consultation.
Implementation reference: skills/_shared/parallel-aware-preamble.md § Implementation.
AUQ reference: skills/_shared/parallel-aware-auq.md.
Phase 1: Plan Verification
Read back the session plan that was agreed at the start. For EACH planned item:
1.1 Done Items
- Verify with evidence: read the changed files, check git diff, run relevant test
- Confirm acceptance criteria are met
- Mark as completed
1.2 Partially Done Items
- Document what was completed and what remains
- Create a VCS issue for the remaining work with:
- Title:
[Carryover] <original task description> - Labels:
priority:<original>,status:ready - Description: what's done, what's left, context for next session
- Title:
- Link to original issue if applicable
1.3 Not Started Items
- Document WHY (blocked? de-scoped? out of time?)
- If still relevant: ensure original issue remains
status:ready - If no longer relevant: close with comment explaining why
1.4 Emergent Work
- Tasks that were NOT in the plan but were done (fixes, discoveries)
- Document and attribute to relevant issues
- If new issues were identified: create them on the VCS platform
1.5 Discovery Scan (if enabled)
Read skills/session-end/discovery-scan.md for embedded discovery dispatch and findings triage.
1.6 Safety Review
Skip if
persistenceisfalsein Session Config (STATE.md won't exist).
Review safety metrics from the session. This is informational — it does NOT block the session close.
Read
<state-dir>/STATE.mdto extract:- Circuit breaker activations: agents that hit maxTurns (
PARTIAL), agents that spiraled (SPIRAL), agents that failed (FAILED) - Worktree status: which agents used worktree isolation, any fallbacks or merge conflicts
- Circuit breaker activations: agents that hit maxTurns (
Read enforcement hook logs from stderr (if captured): count of scope violations blocked/warned, command violations blocked/warned
Summarize:
Safety review: - Agents: [X] complete, [Y] partial (hit turn limit), [Z] spiral/failed - Enforcement: [N] scope violations, [M] command blocks - Isolation: [K] agents in worktrees, [J] fallbacksIf any agents were
SPIRALorFAILED, ensure carryover issues exist (cross-reference with Phase 1.2)Carryover validation fallback (#261): Walk each Wave History entry in STATE.md. For every agent whose status is
SPIRALorFAILED, check whether the line ends with a→ issue #NNNsuffix (or→ existing #NNN). If the suffix is absent, the auto-create call in wave-executor did not run (e.g. a consumer-project #251 V0.x.y-close incident where the session crashed before dispatch completed, or the CLI was offline at detection time). Retroactively file the carryover viacreateSpiralCarryoverIssue:import { createSpiralCarryoverIssue } from '${PLUGIN_ROOT}/scripts/lib/spiral-carryover.mjs'; // For each SPIRAL/FAILED agent missing the "→ issue #NNN" suffix: const result = await createSpiralCarryoverIssue({ taskDescription: '<agent task from Wave History>', kind: 'SPIRAL', // or 'FAILED' context: '<Deviations / error context from STATE.md>', priority: 'high', vcs: '<from Session Config>' }); // result.created → note new issue id in Final Report under "New Issues Created" // result.skipped === 'duplicate' → an earlier session already filed one; record the existing id // result.skipped === 'error' → log in Final Report as "⚠ carryover filing failed for <task>: <error>" and continue (do NOT block close)The module is idempotent via its task-hash dedup marker, so re-running the fallback across sessions will not create duplicates.
1.7 Metrics Collection
Read skills/session-end/metrics-collection.md for JSONL schema and conditional field rules.
1.8 Session Review
Dispatch the session-reviewer agent to verify implementation quality before the quality gate:
On Codex CLI, dispatch via the
session-revieweragent role defined in.codex-plugin/agents/session-reviewer.toml.
- Invoke
subagent_type: "session-orchestrator:session-reviewer"with:- Scope: all files changed this session (from
git diff --name-onlyagainst the base branch) - Context: the session plan (issues, acceptance criteria) and all wave results from STATE.md
- Scope: all files changed this session (from
- Wait for the reviewer's Verdict:
PROCEED — continue to Phase 2
FIX REQUIRED — disposition each listed item by severity:
Finding class Disposition HIGH+ / blocking review finding Fix inline if quick (<2 min); else create an issue ( priority:high,status:ready) and note it in the Final ReportMED / LOW review finding Fold in-session if quick; else record under "Unresolved Review Findings" in the Final Report — DO NOT create an issue (#617) Planned-carryover (item was in the plan, not finished) ALWAYS create a [Carryover]issue per Phase 1.2 — unchangedSPIRAL / FAILED agent carryover ALWAYS file via createSpiralCarryoverIssueper Phase 1.6 — unchanged
1.9 Mission-Status Classification (when mission-status present in STATE.md)
Skip if
persistenceisfalsein Session Config, or ifmission-status:is absent from STATE.md frontmatter. When absent, fall back to binary checkbox detection in 1.1–1.4 unchanged — full backward compat.
When STATE.md frontmatter contains a mission-status: array (set by session-plan + wave-executor per #340), use the enum values to classify items into the 1.1–1.4 buckets. Read the array via parseMissionStatus(frontmatter) from scripts/lib/state-md.mjs.
Classification mapping:
status: completed→ 1.1 Done Items (item finished; verify with evidence per 1.1)status: testingorstatus: in-dev→ 1.2 Partially Done (carryover; document what remains)status: validatedorstatus: brainstormed→ 1.3 Not Started (carryover; check if still relevant)- Items NOT present in the
mission-status:array → fall back to binary checkbox detection per 1.1–1.4 unchanged
Backward compat: When mission-status: is absent from STATE.md (pre-#340 STATE.md files, or sessions where session-plan did not emit the block), behave exactly as before — enum classification is skipped entirely and 1.1–1.4 binary checkbox logic runs as the sole classification mechanism.
1.10 Mission Status Breakdown (when mission-status present)
Skip if
mission-status:is absent from STATE.md frontmatter (backward compat — no breakdown emitted).
After classifying items in Phase 1.9, produce a Mission Status breakdown subsection as part of the closed/carryover summary output. Count the number of tasks at each enum value across ALL waves:
### Mission Status Breakdown
- completed: <N> tasks
- testing: <N> tasks
- in-dev: <N> tasks
- validated: <N> tasks
- brainstormed: <N> tasks
- Total: <N> tasks across <W> waves
Rules:
- Count each task-id entry from the
mission-status:frontmatter array by its currentstatusvalue. completedmaps to Phase 1.1 (Done).testing+in-devmap to Phase 1.2 (Partial).validated+brainstormedmap to Phase 1.3 (Not Started).- Include this block in the Phase 6 Final Report under
### Carried Overor as a standalone subsection immediately after the Completed/Carried Over/New Issues lists. - When all tasks are
completed, the breakdown still appears (confirms clean session state).
Phase 2: Quality Gate
Verification Reference: See
verification-checklist.mdin this skill directory for the full quality gate checklist.
Run ALL checks listed in the verification checklist. If any check fails: fix if quick (<2 min), otherwise create a priority:high issue. Do NOT commit broken code.
Phase 2.0a: Echo-Stub Detection (GH #42)
gate-full.mjs emits a top-level stubbed: {} map in its JSON result, keyed by check name (typecheck, test, lint); value is { kind: 'echo'|'noop' }. When any check was short-circuited as a stub, runCheck() already returned status: 'pass' — so the overall gate verdict is green, but the result is meaningless.
Detection: immediately after parsing the gate-full JSON result, evaluate:
const stubbedEntries = Object.entries(result.stubbed ?? {});
If stubbedEntries.length > 0, surface a HIGH WARN block in the close summary:
⚠ QUALITY GATE STUBBED — <N> command(s) are echo/noop stubs, not real checks:
- <check-name>: <kind> stub (configured: "<command string>")
Re-configure with a real test command in CLAUDE.md Session Config before /close,
OR document this exception in /close --reason.
Behavior by enforcement mode:
enforcement: strict— block /close. Treat as a Phase 2 failure. Present the WARN block and exit without committing.enforcement: warn(default) — continue, but writequality-gate-stubbed: trueto STATE.md Deviations so the metrics writer captures it.enforcement: off— silent. Emit a single-linestderrlog only (echo-stub detected: <check-name>).
Recipe: for container-based test runners (e.g. EspoCRM PHPUnit) where an echo-stub was the historical workaround, see docs/recipes/quality-gate-container-pattern.md.
Source issue: GH #42 (root cause: a consumer-project #251 V0.15.7-close incident — silent false-positive close-verdicts from echo-stub test commands).
2.1 Vault Validation (if configured)
Read skills/session-end/vault-operations.md for validator bash contract and reporting matrix.
2.2 CLAUDE.md (or AGENTS.md) Drift Check (if configured)
Read skills/session-end/drift-operations.md for checker bash contract and reporting matrix. Complements 2.1: vault-sync validates frontmatter inside the vault tree; drift-check validates narrative claims (paths, counts, issue refs, session-file refs) in top-level repo docs.
2.3 Vault Staleness Check (if configured)
Skip this subsection if
vault-staleness.enabledis nottrue(default:false).
Step 1 — Resolve mode
Read vault-staleness.mode from $CONFIG (default: warn). Valid values: off | warn | strict.
If mode === 'off', skip Phase 2.3 entirely.
Step 2 — Invoke staleness probes
Both probes already ship in skills/discovery/probes/. Invoke each via Node import (no shell-out):
import { runProbe as runStaleness } from '$REPO_ROOT/skills/discovery/probes/vault-staleness.mjs';
import { runProbe as runNarrative } from '$REPO_ROOT/skills/discovery/probes/vault-narrative-staleness.mjs';
const projectStaleness = await runStaleness(projectRoot, config);
const narrativeStaleness = await runNarrative(projectRoot, config);
Each probe returns { findings: Array, metrics: Object, duration_ms: Number } and auto-appends a JSONL summary record to its respective metrics file.
Step 3 — Aggregate and route by mode
totalFindings = projectStaleness.findings.length + narrativeStaleness.findings.length
mode === 'warn'(default): report findings to closing report Docs Health line. Never block close.mode === 'strict':- If
totalFindings === 0: continue, logVault staleness: clean (mode=strict). - If
totalFindings > 0: BLOCK the close. Present the findings list and offer override:- On Claude Code: AskUserQuestion with options:
- "Fix and retry Phase 2.3" (Recommended) — exit close, let user investigate
- "Override and close" — proceed, log a Deviation entry in STATE.md
## Deviations:- [<ISO timestamp>] Phase 2.3: Vault staleness strict-mode findings overridden by user. Findings: <count> (projects: <N>, narratives: <M>). - "Abort close" — exit close without writing
- On Codex CLI / Cursor IDE: same options as numbered Markdown list.
- On Claude Code: AskUserQuestion with options:
- If
Step 4 — Surface to closing report
Pass the aggregated counts and mode forward to Phase 6 Final Report (Docs Health line — see Phase 6 below).
Phase 3: Documentation Updates
Final heartbeat (#590-3) — at Phase 3 entry, refresh the session-lock heartbeat BEFORE the multi-minute close-out chain (vault-mirror, dialectic, durable-commit, metrics). A long-idle deep session may not have had PostToolBatch activity for >4h; without a refresh the 4h-TTL lock would lapse mid-close and appear stale to a concurrent session. Place this call BEFORE Phase 3.8 Session Lock Release (which deletes the lock — refreshing a deleted lock is a no-op). Best-effort: a failure must NOT block the close.
// Final heartbeat (#590-3) — refresh before the multi-minute close-out (vault-mirror, dialectic, durable-commit) // so a long-idle deep session's 4h-TTL lock does not lapse mid-close. // BEFORE Phase 3.8 lock-release (which deletes the lock). import { updateHeartbeat } from 'scripts/lib/session-lock.mjs'; updateHeartbeat({ sessionId, repoRoot: process.cwd() });Skip silently if
persistence: falsein Session Config (no session.lock exists in that mode).
3.0 Defensive Cleanup
Delete <state-dir>/wave-scope.json if it still exists:
rm -f <state-dir>/wave-scope.json
This should have been cleaned up by wave-executor after the final wave, but crashed sessions or interrupted executions may leave it behind. A stale scope manifest from a previous session could incorrectly restrict the next session's enforcement hooks.
3.1 SSOT Files
- Update
STATUS.md/STATE.mdif they exist (metrics, dates, status) - Update
CLAUDE.md(orAGENTS.mdon Codex CLI) if patterns or conventions changed during this session - Check
<state-dir>/rules/— if a new pattern was established, suggest a new rule file
3.2 Docs Verification (docs-orchestrator integration)
Skip this subsection if
docs-orchestrator.enabledconfig is nottrue(default:false). Also skip entirely ifdocs-orchestrator.modeisoff.
Reads docs-tasks from STATE.md frontmatter (written by wave-executor Pre-Wave 1b), computes CHANGED_FILES via git diff --name-only "$SESSION_START_REF..HEAD", and runs a per-task verification loop (outcome: ok/partial/gap). In warn mode logs results non-blocking; in strict mode blocks on any gap and presents an AskUserQuestion override prompt. Emits a ### Documentation Coverage (docs-orchestrator) block for inclusion in the Phase 6 Final Report.
See phase-3-2-docs-verification.md for full details.
3.2a Session Handover (for significant sessions)
If this session made substantial changes, create or update:
<state-dir>/session-handover/doc with: tasks completed, resume point, metrics changed, issues opened/closed- Or update
<state-dir>/STATE.mdwith session digest
3.3 Claude Rules Freshness
Review <state-dir>/rules/ files that are relevant to this session's work:
- Are the rules still accurate after this session's changes?
- Should any rule be updated with new patterns?
- Should a new path-scoped rule be created?
- Suggest changes but DO NOT modify without user confirmation
3.4 Update STATE.md
Ownership Reference: See
skills/_shared/state-ownership.md. session-end is authorized to setstatus: completedplus the optionalupdatedtimestamp (#184), and — as of Phase A of Epic #271 — the 5 Recommendation fields written by Phase 3.7a. No other fields.
Runtime Ordering Note (Epic #271 Phase A): Phase 3.4's
status: completedwrite executes LAST in Phase 3, AFTER Phase 3.7 (sessions.jsonl) and Phase 3.7a (Compute and Write Recommendations). The ordinal position here (3.4) is kept for historical compatibility; the canonical runtime order is3.1 → 3.2 → 3.3 → 3.4a → 3.5 → 3.5a → 3.6 → 3.6.5 → 3.6.7 → 3.7 → 3.7a → 3.7b → 3.4. Rationale: Phase 3.7a reads in-memory session metrics and writes the 5 Recommendation fields viaupdateFrontmatterFields; that write must complete BEFORE the STATE.md frontmatter is finalized withstatus: completedso the Recommendation fields are visible to the next session-start while STATE.md is stillstatus: active. Crash-resilience: if/closeaborts between 3.7a and 3.4, STATE.md carriesstatus: active+ Recommendations; session-start Phase 1.5 offers resume (and the banner renders). If the reverse ordering were used (status: completed first), a crash would leavestatus: completedwithout Recommendations — the Reader would silently no-op the banner, losing the handoff.
Gate: Only run if
persistenceis enabled in Session Config and<state-dir>/STATE.mdexists.
- Set frontmatter
status: completed - Record final wave count and completion time in the frontmatter
- Touch
updated: <ISO 8601 UTC>in the frontmatter (issue #184). Usescripts/lib/state-md.mjs→touchUpdatedFieldfor safety:
Silent no-op if the file has no frontmatter.node --input-type=module -e " import {readFileSync, writeFileSync} from 'node:fs'; import {touchUpdatedField} from '${PLUGIN_ROOT}/scripts/lib/state-md.mjs'; const p = '<state-dir>/STATE.md'; writeFileSync(p, touchUpdatedField(readFileSync(p, 'utf8'), new Date().toISOString())); " - Keep the file as a record — do NOT delete it (next session-start reads it)
If STATE.md doesn't exist, skip this subsection.
3.4a Coordinator Snapshot Cleanup (#196)
Pre-dispatch snapshots (refs/so-snapshots/<sessionId>/wave-*) are created by wave-executor before each wave dispatch so that session-start can offer recovery if a session is interrupted mid-wave. On a clean close those snapshots are no longer needed and should be deleted. In addition, orphaned refs from older sessions that were never cleaned up (e.g. after a hard crash) are garbage-collected using an age-based policy (14 days).
Gate: Only run if
persistenceistruein Session Config. Skip entirely when persistence is off (snapshots are never written in that mode).
node --input-type=module -e "
import { listSnapshots, deleteSnapshot, gcSnapshots } from '${PLUGIN_ROOT}/scripts/lib/coordinator-snapshot.mjs';
// Step A: delete this session's snapshots (clean close → we don't need them)
const mine = await listSnapshots({ sessionId: '${SESSION_ID}' });
for (const s of mine) {
const r = await deleteSnapshot({ refName: s.ref });
if (!r.ok) console.error('snapshot cleanup:', r.error);
}
// Step B: GC orphans older than 14 days (non-fatal)
const gc = await gcSnapshots({ olderThanDays: 14 });
console.log(\`snapshot cleanup: deleted \${mine.length} from this session + \${gc.deletedCount} expired orphans (scanned \${gc.scanned}).\`);
"
Failures in either step are logged to stderr but do not block session close — a missed cleanup is self-healing via the 14-day GC on the next session.
This cleanup is the counterpart to the session-start Phase 1.5 recovery prompt: once a session closes cleanly, future sessions must not be offered recovery for its snapshots.
3.5 Session Memory
Gate: Only run if
persistenceis enabled in Session Config AND platform is Claude Code (session memory at~/.claude/projects/is Claude Code-only). Learnings (Phase 3.5a) and metrics (Phase 3.7) still write to.orchestrator/metrics/on all platforms.
- Create
~/.claude/projects/<project>/memory/session-<YYYY-MM-DD>.mdwith:- Frontmatter:
name,description(1-line summary),type: project ## Outcomes— per-issue status (completed / partial / not started) with evidence## Learnings— patterns discovered, architectural insights, gotchas## Next Session— priority recommendations, suggested session type, blockers
- Frontmatter:
- Update
~/.claude/projects/<project>/memory/MEMORY.md:- Under a
## Sessionsheading (create if missing), add:- [Session <date>](session-<date>.md) — <one-line summary>
- Under a
3.5a Learning Extraction + 3.6 Memory Cleanup & Learnings Write
Read skills/session-end/learning-patterns.md for extraction heuristics, confidence updates, passive decay, and JSONL write procedure.
3.6.3 Memory Proposals Collection (#501, F2.1)
Gate: Skip this phase entirely when ANY of:
persistenceisfalsein Session Configmemory.proposals.enabledisfalse(default:true).orchestrator/metrics/proposals.jsonldoes not exist OR contains zero entries
After learnings are written (Phase 3.6) and BEFORE auto-dream dispatch (Phase 3.6.5), collect agent-proposed memory entries written during this session and present them to the operator via AskUserQuestion multiSelect. Approved entries flow to learnings.jsonl with _provenance: agent-proposed@<wave-id>. Rejected entries are archived to .orchestrator/proposals.rejected.log.
The proposals queue is populated mid-session by wave-executor agents calling node scripts/memory-propose.mjs --type ... --subject ... --insight ... --evidence ... --confidence .... The CLI enforces:
- Quota per wave (default 5, configurable via
memory.proposals.quota-per-wave) - Confidence floor (default 0.5, configurable via
memory.proposals.confidence-floor) - Wrong-context guard (CLI exits non-zero when STATE.md
statusis notactive)
Coordinator-direct procedure
Read Session Config:
memory.proposals.enabled(defaulttrue),memory.proposals.quota-per-wave(default 5),memory.proposals.confidence-floor(default 0.5),auto-dream.min-confidence(default 0.5 — issue #566; SECOND gate above the write-timememory.proposals.confidence-floor).Invoke
collectProposalsfromscripts/lib/memory-proposals/collector.mjs, passing the collect-emit confidence floor from Session Config:import { collectProposals } from '${PLUGIN_ROOT}/scripts/lib/memory-proposals/collector.mjs'; const { queue, stats, perWaveSummaries } = await collectProposals({ repoRoot: process.cwd(), // Issue #566: collect-emit confidence floor. Records with // `record.confidence < minConfidence` are dropped from `queue` (but // counted in stats). When the key is absent, defaults to 0.5 via the // `_parseAutoDream` parser. minConfidence: config['auto-dream']?.['min-confidence'], });If
queue.length === 0: logmemory-proposals: queue empty (stats: ${JSON.stringify(stats)})and continue.AUQ pagination logic: partition the queue into FIFO batches of 4 inline:
- Empty queue → silent skip (no AUQ rendered).
- 1-4 items → single multiSelect call with all items as options.
- 5+ items → sequential multiSelect calls in batches of 4 (FIFO order; final batch may have < 4 items).
// Inlined from former scripts/lib/memory-proposals/auq-partition.mjs (PRD F2.2 #502 closed; see #558 M2). const BATCH_SIZE = 4; const batches = []; if (Array.isArray(queue) && queue.length > 0) { for (let i = 0; i < queue.length; i += BATCH_SIZE) { batches.push(queue.slice(i, i + BATCH_SIZE)); } }Then iterate
batchesand emit oneAskUserQuestionper batch withheader: "Memory — Confirm Proposals (Batch N of M)". Option label format:[<type-12>] | <subject-40> | conf=X.XX. Option description:evidence: <first 60 chars of insight>.multiSelect: true.After all batches answered, partition the queue into
approved(any option selected across all batches) andrejected(all unselected).Invoke
writeApprovedandarchiveRejectedfromscripts/lib/memory-proposals/sink.mjs:import { writeApproved, archiveRejected, clearProposalsJsonl } from '${PLUGIN_ROOT}/scripts/lib/memory-proposals/sink.mjs'; const writeResult = await writeApproved({ approved, repoRoot, sessionId }); const archiveResult = await archiveRejected({ rejected, repoRoot, reason: 'user-declined' }); await clearProposalsJsonl({ repoRoot });Log outcome for Phase 6 Final Report:
memory.proposals: <queued> queued → <approved> approved, <rejected> rejected (dropped: <dropped> quota, <below_floor> below-floor).
Failure modes
- If
collectProposalsfails (fs error): log warning⚠ memory-proposals: collect failed (${err}) — skipping, do not block session close. - If
writeApprovedreports errors per-record: log each, but continue (per-record fault isolation per sink contract). - If
clearProposalsJsonlfails: log warning; do not block. The file may be re-collected at the next session-end, idempotent.
Cross-references
- PRD:
docs/prd/2026-05-21-learning-memory-modernization.md§ F2.1 - Modules:
scripts/lib/memory-proposals/{schema,store,collector,sink}.mjs - CLI:
scripts/memory-propose.mjs(agents call this) - Hook:
hooks/pre-bash-memory-propose-audit.mjs(audit trail) - Coordinator AUQ spec:
agents/memory-proposal-collector.md(reference doc) - Sibling phases: 3.6.5 Auto-Dream (#502), 3.6.7 Auto-Dialectic (#506)
- Issue: #501
3.6.5 Auto-Dream Dispatch (#502, F2.2)
Skip this phase if
memory-cleanup-threshold: 0(kill-switch per PRD F2.2). Also skip on non-Claude-Code platforms (memory dir at~/.claude/projects/is Claude Code-only, mirrors Phase 3.5 gate).
After learnings are written (Phase 3.6), determine whether to emit a manual-cadence nudge to run /memory-cleanup --dry-run in the next session. The decision uses MEMORY.md line count and a sessions-since-last-cleanup signal. There is no memory-cleanup agent in the registry, so the historical auto-dream subagent dispatch never fired (see #614) — the nudge replaces it. A manually-run /memory-cleanup --dry-run writes a unified-diff proposal to .orchestrator/pending-dream.md for the session after that to apply via /memory-cleanup --apply-pending.
Read
memory-cleanup-threshold(default 5) andmemory-cleanup-soft-limit(default 180) from$CONFIG.Invoke
shouldDispatchAutoDreamfromscripts/lib/auto-dream.mjs:import { shouldDispatchAutoDream } from '${PLUGIN_ROOT}/scripts/lib/auto-dream.mjs'; import { resolveMemoryDir } from '${PLUGIN_ROOT}/scripts/lib/memory-paths.mjs'; const memoryDir = resolveMemoryDir(); const decision = await shouldDispatchAutoDream({ repoRoot: process.cwd(), memoryDir, threshold: config['memory-cleanup-threshold'] ?? 5, softLimit: config['memory-cleanup-soft-limit'] ?? 180, });If
decision.trigger === false: logauto-dream: not triggered (${decision.reason})and continue. Emit no nudge.If
decision.trigger === true: do not dispatch a subagent — there is nomemory-cleanupagent inagents/, so the historicalAgent({…})dispatch pointed at the agent namememory-cleanup(a subagent type that was never built) and never fired (see #614). Instead, emit a manual-cadence nudge and continue:auto-dream: cadence reached (${decision.reason}) — run /memory-cleanup --dry-run manually in the next session, then apply the proposal with /memory-cleanup --apply-pending.The
shouldDispatchAutoDreamdecision helper andscripts/lib/auto-dream.mjslib stay in use: they compute the signal that drives this nudge and back the manual/memory-cleanuppath (writePendingDream/readPendingDream/applyPendingDream).Record the outcome (skipped / nudge-emitted) so Phase 6 Final Report can surface a line:
auto-dream: manual /memory-cleanup --dry-run recommended (cadence reached) — apply with /memory-cleanup --apply-pending next session.
The pending-dream sidecar at .orchestrator/pending-dream.md is intentionally outside the vault tree — vault-mirror (Phase 3.7) must exclude it from its scope so the proposal survives the session close without being mirrored into 50-sessions/.
Cross-reference: PRD F2.2 acceptance criteria; scripts/lib/auto-dream.mjs API (shouldDispatchAutoDream, readDreamSignals, writePendingDream, readPendingDream, applyPendingDream).
3.6.7 Auto-Dialectic Dispatch (#506, F2.5)
Skip this phase if
dialectic.cadence: 0(kill-switch per PRD F2.5 AC3). Also skip ifpersistenceisfalsein Session Config.
After learnings are written (Phase 3.6) and the auto-dream decision is made (Phase 3.6.5), determine whether to emit a manual-cadence nudge to run /evolve --dialectic in the next session. The decision uses sessions-since-last-dialectic counted against .orchestrator/dialectic-last-run. There is no evolve agent in the registry, and the nearest one (dialectic-deriver) is sandbox-tier: read-only and cannot write the sidecar — so the historical auto-dialectic subagent dispatch never fired (see #614). On trigger, emit the nudge and advance .orchestrator/dialectic-last-run; the timestamp is updated only when the nudge is emitted (not on skip), so the reminder surfaces once per cadence window rather than every session. A manually-run /evolve --dialectic --dry-run writes the proposed diff to .orchestrator/dialectic-pending.md.
Read
dialectic.cadence(default 5),dialectic.model(default haiku),dialectic.budget-tokens(default 8000) from$CONFIG.Invoke
shouldDispatchAutoDialecticfromscripts/lib/auto-dialectic.mjs:import { shouldDispatchAutoDialectic } from '${PLUGIN_ROOT}/scripts/lib/auto-dialectic.mjs'; const decision = await shouldDispatchAutoDialectic({ repoRoot: process.cwd(), cadence: config.dialectic?.cadence ?? 5, });If
decision.trigger === false: logauto-dialectic: not triggered (${decision.reason})and continue. Emit no nudge. Do NOT update.orchestrator/dialectic-last-run.AC4 precondition guard: Even if cadence met, if
signals.sessionsSinceLast === 0 && signals.learningsSinceLast === 0, skip with reasonno-new-input-since-last-run. The Final Report (Phase 6) MUST include the literal stringdialectic: skipped (no new input since last run).If
decision.trigger === true: do not dispatch a subagent (see #614 — noevolveagent exists;dialectic-deriveris read-only and cannot write the sidecar). Instead, emit a manual-cadence nudge and continue:auto-dialectic: cadence reached (${decision.reason}) — run /evolve --dialectic --dry-run manually in the next session, review .orchestrator/dialectic-pending.md, then apply with /evolve --dialectic --apply.The
shouldDispatchAutoDialecticdecision helper andscripts/lib/auto-dialectic.mjslib stay in use: they compute the cadence signal that drives this nudge.When the nudge is emitted (cadence reached), update
.orchestrator/dialectic-last-runviawriteDialecticLastRun({ repoRoot, isoTimestamp: new Date().toISOString() })so the cadence counter advances and the nudge does not repeat every session. Atomic; failures non-fatal.Record outcome (skipped / nudge-emitted) for Phase 6 Final Report:
auto-dialectic: manual /evolve --dialectic --dry-run recommended (cadence reached) — apply with /evolve --dialectic --apply next session.
The .orchestrator/dialectic-pending.md sidecar is intentionally outside the vault tree — vault-mirror (Phase 3.7) MUST exclude it from its scope.
Cross-reference: PRD F2.5 acceptance criteria (#506); scripts/lib/auto-dialectic.mjs API.
Dialectic chain rationale — design choices in the manual
/evolve --dialecticchain (/evolve → runDialecticDeriver → dispatchAgent → Agent). Session-end no longer auto-dispatches this chain (see #614 — theevolveagent never existed); the rationale below applies when you run/evolve --dialecticmanually:
- /evolve → subagent (not direct invoke): the manual
/evolve --dialecticskill spawns a subagent so the dialectic pass runs in a fresh context window — keeping the deriver's input-heavy payload (top-50 learnings + last-10 sessions + 2 peer cards + steering) out of the invoking coordinator's context, and letting the deriver run as Haiku while the coordinator stays Opus.- /evolve → runDialecticDeriver (not direct dispatchAgent): /evolve owns argument parsing, config resolution, dry-run/apply gating, error-handling, and sidecar writes; runDialecticDeriver owns the pure derivation pipeline (load → payload → budget-check → dispatch → parse → guard). Separating skill-level orchestration from deriver business logic lets unit tests exercise the deriver without standing up the full evolve skill.
- runDialecticDeriver → dispatchAgent (DI boundary): per
.claude/rules/prompt-caching.md:3, session-orchestrator forbids direct@anthropic-ai/sdkimports in business logic (the harness manages caching at the platform layer). dispatchAgent is the injected boundary — the evolve skill wires the realAgent({...})harness call at runtime, tests pass avi.fn()mock. Same DI shape asscripts/lib/autopilot.mjs::runLoop({opts})(cf.scripts/dialectic-deriver.mjs:7-16,531).
3.7 Write Session Metrics
Read skills/session-end/session-metrics-write.md for JSONL append, vault-mirror invocation, and behavior matrix.
3.7a Compute and Write Recommendations (Epic #271 Phase A)
Gate: Only run if
persistenceistruein Session Config AND<state-dir>/STATE.mdexists. Skip silently otherwise.
Ownership Reference: See
skills/_shared/state-ownership.md. session-end is the ONLY writer of the 5 Recommendation fields (recommended-mode,top-priorities,carryover-ratio,completion-rate,rationale). No other skill may write these keys.
Ordering: Runs AFTER Phase 3.7 (sessions.jsonl is just-written — reads in-memory session metrics, NOT JSONL) and BEFORE Phase 3.4
status: completedsetting. See the Phase 3.4 Runtime Ordering Note for rationale.
Calls computeV0Recommendation({completionRate, carryoverRatio, carryoverIssues}) from in-memory session metrics and writes 5 fields to STATE.md frontmatter via updateFrontmatterFields. Inputs MUST come from in-memory metrics, NOT re-read from sessions.jsonl. On any exception writes recommendation-compute-failed to sweep.log and does NOT block Phase 3.4.
See phase-3-7a-recommendations.md for full details.
3.7b Durable-Commit Session Telemetry (#490 AC2)
Gate: Always runs when persistence is enabled. Local execution is a no-op (
enabled: false).
Ordering: Runs AFTER Phase 3.7a (Recommendations written to STATE.md) and BEFORE Phase 3.4 (
status: completed). See the Phase 3.4 Runtime Ordering Note canonical order.
Wraps the already-completed Phase 3.7 + 3.7a writes with withDurableCommit (from scripts/lib/autopilot/durable-telemetry.mjs) for the two session-end-owned files: .orchestrator/metrics/sessions.jsonl and <state-dir>/STATE.md. enabled: false keeps local closes a no-op ({ok: true, skipped: true}); the flag flips true only in cloud Routines execution so telemetry survives ephemeral-clone reclamation. autopilot.jsonl is NOT in scope here — scripts/lib/autopilot/loop.mjs owns its commit (#490 Wave-2).
See phase-3-7a-recommendations.md § Phase 3.7b for the full withDurableCommit invocation.
Phase 3.8: Session Lock Release (#330)
Gate: Only run if
persistenceistruein Session Config. Skip silently otherwise.
After STATE.md is finalized with status: completed (Phase 3.4) and Recommendations are written (Phase 3.7a), release the distributed session-lock so the next session can acquire it cleanly:
import { release } from 'scripts/lib/session-lock.mjs';
// sessionId = the session identifier established by session-start Phase 1.2 acquire()
// and stored in .orchestrator/session.lock (session_id field); matches the
// STATE.md frontmatter `session:` field written during Pre-Wave 1b initialization.
const result = release({ sessionId, repoRoot: process.cwd() });
// result.ok is always true unless a filesystem error occurred.
// result.deleted === true → lock file removed successfully.
// result.deleted === false → lock was absent or belonged to a different session_id (silent-OK).
If result.deleted === false, log info: session-lock not released — already absent or session_id mismatch (no action needed) and continue. This is a non-error state.
If result.ok === false (rare filesystem error), log ⚠ session-lock: release failed — <result.reason> and continue. Do NOT block the close for a lock-release failure — the TTL provides automatic expiry for the next session.
The lock is released here — AFTER all STATE.md writes are complete and BEFORE the commit is staged in Phase 4.1. This ordering ensures a clean handover: the lock file is absent from the working tree when the commit is assembled, so it is not accidentally staged.
Phase 4: Commit & Push
4.1 Stage Changes
- Stage files individually:
git add <file>— NEVERgit add .orgit add -A - Always stage these session artifacts (if modified):
.orchestrator/metrics/sessions.jsonl(session summary from Phase 3.7).orchestrator/metrics/learnings.jsonl(learnings from Phase 3.6)<state-dir>/STATE.md(session state, if persistence enabled)- Any files created or modified by wave agents
- Review staged changes:
git diff --cached— verify every change is from THIS session - If you see changes you did NOT make, ask the user (parallel session awareness)
4.2 Commit
Use Conventional Commits format:
type(scope): description
- [bullet points of what changed]
- Closes #IID1, #IID2 (if applicable)
Co-Authored-By: Claude Opus 4.6 (1M context) <noreply@anthropic.com>
For sessions with many changes, prefer ONE commit per logical unit (not one mega-commit).
4.3 Push
git push origin HEAD
4.4 GitHub Mirror (if configured in Session Config)
# Only attempt if 'mirror: github' is in Session Config AND remote exists
git remote get-url github 2>/dev/null && git push github HEAD 2>/dev/null || echo "GitHub mirror: not configured"
Phase 4a: Auto-Promoted Worktree Cleanup (#575 P3.2)
Skip if
persistence: falsein Session Config. Skip silently if the current worktree is NOT an Auto-promoted sibling (the common case).
After Phase 4 commit+push has durably persisted `sessions.j
…(truncated)