Skip to content

Measure the chain in real use, and revisit its defaults #12

Description

@PierreMardon

Status

Open, not done: a way to start without waiting for real use is proposed. On 2026-10-08 the developer proposed to clone a public project with a strong harness, to be selected, and to play real needs there, a whole campaign on the prompts as they stand among them: see the comment of that day. Read again on 2026-10-05: the defaults named here still stand, max_autonomous_passes is 3, and since #40 the same ceiling bounds the reworks in planning; surface-status gate runs the project's full gates; every slice, review and fix gets a fresh agent.

What to measure

The chain's defaults were set before any real use: three autonomous passes, full gates run by the script at every pass, a fresh agent for every slice, review and fix. Once projects use it, measure per plan:

  • the passes used, and the stops by cause (conformant, ceiling, break, alarm, refusal);
  • the review findings by class (defect, deviation, break), and the breaks the developer accepted or refused;
  • the gate runs, their duration and their share of the loop's time;
  • the cost and the duration of each command.

Most of it can be read from the journals and reports the plans leave on the main branch; a surface-status subcommand could aggregate them.

Then revisit the defaults with the numbers: the passes, the gate runs at every pass (their commands are found at planning since #18), the fresh agent per step.

Activity

Sign up for free to join this conversation on GitHub. Already have an account? Sign in to comment

Metadata

Metadata

Assignees

No one assigned

    Labels

    No labels
    No labels

    Type

    No type

    Projects

    No projects

      Milestone

      No milestone

      Relationships

      None yet

      Development

      No branches or pull requests

      Issue actions