Repository navigation
docs: phased build-out plan for the research loop (substrate, depth, parallel) - #127
Conversation
…parallel) A reviewable plan that sequences the three efforts toward research-loop.md's north-star: (1) consolidate the role-runner (thin, feature-driven), (2) the depth axis (k as a contract dial, generalizing panel-on-wake), (3) parallel dispatch (portfolio + deadline-salvage), with the benchmark-integrity gate threaded through as the governor and the dispatcher-must-go-live dependency called out. Cross-linked from research-loop.md. Co-Authored-By: Claude Opus 4.8 <noreply@anthropic.com>
There was a problem hiding this comment.
Round 1 — reviewed head 55937afd — reviewer hermes/gpt-5.6-terra.
second opinion — terra
Advisory findings from autoresearch — the code owner decides. Reply to disagree; the autoresearch:no-review label opts this PR out.
Verdict: nothing blocking — 1 advisory note.
1 finding attached to the lines below.
There was a problem hiding this comment.
Round 1 — reviewed head 55937afd — reviewer claude/claude-opus-5.
Advisory findings from autoresearch — the code owner decides. Reply to disagree; the autoresearch:no-review label opts this PR out.
Verdict: nothing blocking — 3 advisory notes.
Advisory (non-blocking):
- Phase 1 premise is stale: the authoring roles already run through
run_role(docs/design/research-loop-buildout.md:30; high) - Phase 1 acceptance criteria are already met by current code (
docs/design/research-loop-buildout.md:46; high) parallel-climbreference does not resolve to anything in the repo (docs/design/research-loop-buildout.md:95; low)
Verified doc cross-references that do resolve: consolidation.md, agent-substrate.md ("The role-runner: one loop replaces five drivers", line 266), dispatcher.md, scaling.md, orchestrator-verify.md, judge-placement.md, plus the code hooks panel_reads/panel_revisions (orchestrator.py) and supports_resume (harness.py, climb.py:777). Claims I could not check from the repo: the dispatcher being "proven but dark" on real Slurm, PR numbers #123/#125, and the climb-lessons notes (the notebook is a separate private repo per architecture.md).
… run_role Review caught that the five-drivers→run_role consolidation is already DONE (consolidation.md; climb/steward/followup all call run_role). Reframe Phase 1 as the layer ABOVE run_role — the shared run→await→decide composition seam that depth and parallel both need — rather than a from-scratch role-runner consolidation. Co-Authored-By: Claude Opus 4.8 <noreply@anthropic.com>
What
A plan design doc (
docs/design/research-loop-buildout.md) that turns theresearch-loop.mdnorth-star into a concrete, phased build-out — the threeefforts we converged on, sequenced with their dependencies.
The shape
One substrate, two axes, one governor:
shared
run_role, as the judges already are).kaxis) — the agent iterating on its own results,ka contractdial, generalizing the panel-on-wake slices (
k = 1today).M-within-a-climb) — a portfolio on one objective,deadline-salvaged.
through so scaling the engine doesn't scale benchmark-gaming.
Sequencing (and why)
dep). Enables the rest; avoids each axis growing its own per-lane
orchestration (the Tick: drive the codex author across the climb-authoring lanes #123 lesson).
cross-session depth and parallel (it's proven but dark today).
taxonomy + suite gate landing alongside.
Depth before parallel: depth extends something already built and is cheaper to
make honest; parallel is the larger new capability.
Status
Status: plan— reviewable before any code. Non-goals and open questions arecalled out in-doc (esp. for parallel: workspace sharing, member interaction,
budget split, salvage ranking, depth×parallel composition). Cross-linked from
research-loop.md.No secrets / no large files
Docs only.
🤖 Generated with Claude Code