Skip to content

Execute the merge driver's printed recipe, and file the class its defect belongs to - #10505

Merged
gunbai-bot[bot] merged 32 commits into
mainfrom
session/wise-bat-147
Sep 5, 2026
Merged

gunbai-bot[bot] merged 32 commits into
mainfrom
session/wise-bat-147

Conversation

@gunbai-bot

@gunbai-bot gunbai-bot Bot commented Sep 5, 2026 •

Copy link
Copy Markdown
Contributor

What this is

A follow-up to #10460, which repaired the merge driver's printed repair recipe. That PR enrolled the evidence it could reach — test.claim.generated_artifact_merge_driver_recipe asserts over the emitted script — and its own header says those reads "cannot claim the command RUNS", naming test.claim.generated_artifact_merge_driver_real_execution as the wet arm that would. That arm carried no recipe claim. This adds it, lifts the extractor so there is one program rather than a copy of one, and files the failure-mode class #10460 held back.

A string read proves an escape was APPLIED; only stderr after expansion proves it is CORRECT

Those are different claims. Doubling the wrong character, or escaping $ and killing the $merged_path the driver defines, both produce escaped bytes and a recipe that still does not work — and both pass an emitted-bytes assert. driver_printed_step_two_carries_the_row_extractor_by_real_execution drives a real refused merge and reads merged.stderr, which is what the reader actually receives.

The reconciliation is three sets, not one number

A cardinality equality is a summary of an identity join, not the join. Two mutations preserve the count and pass it, so both are enrolled as their own controls. Measured against the fixture, all three count-equal:

fixture extracted missing unexpected duplicated
full 157/157 — — —
substitution 157/157 empty_capture_read_as_clean_result not_a_rostered_row_identity —
duplication 157/157 empty_capture_read_as_clean_result — Heal job for generated artifacts

The fixture is built from recurring_failure_mode_roster and rung_drop_roster and carries both row spellings, because the hole #10460 found on docs/design-rung-drops.md was a correct program pointed at a row shape it could not match: it rendered perfectly, read zero of 36 rows, and reported a pass it never measured. Non-emptiness cannot see that — it passes at one row, at forty, and at all-but-one. The plain removal control is enrolled for the consumer consequence (the difference names which rows went dark), and explicitly not as evidence about which join was built, since it passes under both.

The extractor is one declaration

The landed step 2 spells a rows() shell function with two sed -e clauses inside a prose sentence, so the printed program and any executed copy are two things that can drift. They are lifted to generated_artifact_merge_driver_row_identity_extraction_expressions; the printed text is derived from them and the witness executes them. -e pairing lives in extdeps.tools.sed beside the -i it already owned; single-quote wrapping is posix_single_quote.

The emitted bytes are unchanged — git diff origin/main -- .githooks/ is empty.

The red, measured

Reverting the escape gives 2 command not found lines, a printed extractor of s/^- .*/\1/p, and 0 rows from the 157-row fixture — not 36. An invalid backreference makes sed refuse the entire program, so one mangled character silences a clause nothing touched.

The ledger row

printed_remedy_is_mangled_by_the_medium_that_prints_it, identified by calm-owl-417 and held out of #10460 because filing needed roster.dag and the regenerated projection while both were moving — a follow-up that did not land before that session ended, leaving the class only in a merged PR body. Their 7a/7c specimens are carried as theirs, with that provenance on the carrier.

Beyond them the row carries the recognition rule (sender sees intact text, recipient gets a silently shortened sentence, nothing reports a failure); the transport-success invariant (delivery answers whether the submitted payload arrived, never whether it preserved the intended message); the anchoring rule (a preservation digest computed after the transformation faithfully attests the corrupted payload); and two further specimens in agent messages, one of which was a message about this class.

And the mangling executes. A message containing -> Bool inside double quotes ran - as a command and > Bool as a redirection, creating a file in the working tree — a fragment of English prose evaluated as a program. It did not reach a commit because the last commit happened to precede it by fifteen minutes; the row records the ordering as the reason, because recording "it did not land" would be luck reported as safety.

Verification status

All five wet functions return true by execution on a tree-built gunbc, with a must-fail control (NoSuchFunction) proving the corpus typechecks with this branch in it and that evaluation is reached. Details and the control output are in the comment thread.

The shell-layer measurements above are reproduced by those witnesses rather than replacing them.

docs/design-failure-modes.md was regenerated by heal-generated-artifacts (e77236b), which is step 5 of the driver's own recipe. A later merge of main hit the driver on that same projection; it refused as designed (0 conflict markers, 3 index stages), and following its printed recipe verbatim gave an empty step-2 answer — with the ours side shown to have dropped exactly one row at an identical row count, which is the substitution case this PR enrolls, occurring in production.

Also

One real parse defect fixed: extdeps.tools.sed carried its new operation's annotation inside the service block, which §4c forbids. Invisible in a diff, caught only by running the compiler over it.

🤖 Generated with Claude Code

https://claude.ai/code/session_01F8wZ8BQ3mXe4cSfqDJiWNj

Brian Searls and others added 16 commits September 4, 2026 19:23
… echo that prints it, one never matched half its subject

Three defects in gunbc.generated_artifact_merge_driver, all one class: a
documented command that does not do what it says when run verbatim.

1. AuthorRegeneratesBeforeMerge step 3 ended `claim_executor
   --required-regen-fixed-point` with no --source-root. That binary's
   argument guard sits ABOVE the fixed-point branch, so the run exits 2 on
   `provide at least one --source-root` before reaching a mode that reads no
   roots at all -- which is why the omission looked harmless to write. It is
   the last command of the arm every generated artifact takes except the two
   heal-delegated rosters, so the author has paid two builds and two regen
   passes when it refuses.

2. The heal arm's set-difference command is written between backticks, and
   the driver echoes every recipe line inside DOUBLE quotes (they are
   templates; $merged_path must interpolate). Executed against the committed
   .githooks/generated-artifact-merge: bash ran the capture group as a
   command substitution -- two `([a-z0-9_]*): command not found` lines -- and
   printed the instruction as `sed -n 's/^- .*/\1/p'`, capture group deleted.
   Pasted, that extracts zero rows from both sides and the difference over
   two empty sets is empty, which the recipe's own words read as MUST BE
   EMPTY: the mangled remedy reports the vacuous pass the intact remedy
   exists to prevent. Escaping now happens in emit_driver_stderr_echo, where
   the quoting decision is made; it deliberately differs from
   gunbc.shell_bash_runner bash_escape_for_double_quotes on exactly one
   character ($), and that difference is named in the module rather than
   left to read as a fork.

3. The same command's row extractor only ever matched
   docs/design-failure-modes.md. The other delegated projection,
   docs/design-rung-drops.md, carries `### Title — declared ...` headings,
   not slug bullets: 105 rows extracted from one file, 0 from the other, so
   on the rung-drops path the check was green by construction. One extractor
   now names both row identities.

Evidence: test.claim.generated_artifact_merge_driver_recipe reads the
EMITTED script rather than the authored list -- the second defect is
invisible in the authored string -- and refuses the unescaped form.
gunbc.recurring_failure_mode.printed_remedy_is_mangled_by_the_medium_that
_prints_it files the class, with its boundary against
restoration_promise_names_a_route_that_does_not_exist (there the route does
not exist; here it does, and the rendering is what breaks, so every by-name
instrument agrees the remedy is present).

No emitted artifact hand-edited.

Co-Authored-By: Claude Opus 5 <noreply@anthropic.com>
Claude-Session: https://claude.ai/code/session_01BPNfohdhatLvf5HT9jK5k6
…ign-failure-modes.md are contended by three open PRs

The class is real and stands (a printed remedy mangled by the medium that
prints it, whose failure arm is a vacuous PASS rather than a refusal), but
filing it needs a row file AND an import/entry edit in
gunbc.recurring_failure_mode.roster, plus the regenerated
docs/design-failure-modes.md projection -- and #10317, #10320 and #10326 are
open against exactly those two paths. Landing it here would put this branch
into a driver-refused merge on the very projection whose repair recipe this
branch is fixing.

So the driver repair lands alone, touching no contended path. The row
follows in its own change against whatever roster head survives those three.

Co-Authored-By: Claude Opus 5 <noreply@anthropic.com>
Claude-Session: https://claude.ai/code/session_01BPNfohdhatLvf5HT9jK5k6
… into the instruction it printed

Caught by simulating the emitter over the authored strings and RUNNING the
result, which is the only reading that sees this class at all. The recipe
had been drafted as `rows() { git show "$1:$merged_path" | ... }` -- and $
stays active in the emitted echo by construction, because $merged_path must
interpolate. So $1 was git's %O placeholder, and the printed instruction
read `git show "O:docs/design-rung-drops.md"`: an author pasting it would
have asked git for a ref named O. Escaping does not save it either, since
the escaper doubles backslashes before the reader sees them, so `\$1` prints
as a backslash.

The line now names $merged_path -- the one variable the driver defines --
and pipes into a parameterless function. Witnessed by
heal_recipe_names_only_the_variable_the_driver_defines.

EXECUTED, as rendered, against a real divergence (this branch's base vs
origin/main, which is ahead by a landed roster PR):

  failure-modes  base=112  head=105  and the difference NAMES the seven rows,
  rung-drops     base=35   head=35   difference empty

Seven named rows and a measured empty -- neither of which the recipe could
produce before this branch, on either projection.

Co-Authored-By: Claude Opus 5 <noreply@anthropic.com>
Claude-Session: https://claude.ai/code/session_01BPNfohdhatLvf5HT9jK5k6
…is carrier's medium on the carrier

Answers review 60392 (codex/gpt-5.6-sol, REQUEST_CHANGES) on its two
actionable points.

THE SECOND QUOTING IMPLEMENTATION IS GONE, which was the fair half of the
finding: I had authored an escaper beside gunbc.shell_bash_runner
bash_escape_for_double_quotes, and one concept with two names is what DESIGN
section 3 forbids -- the next fix lands in one of them. But the two callers
genuinely disagree about exactly one character. A literal line must have its
$ escaped or the shell expands a variable the author never wrote; a TEMPLATE
line, which is what a recipe naming $merged_path is, must leave $ active or
the instruction prints a dollar sign instead of the path. So the shared
function is SPLIT rather than copied:
bash_escape_double_quote_specials_except_expansion holds the identical part,
bash_escape_for_double_quotes is now composed from it plus the $ arm, and
this module calls the shared arm. Backslash-first is preserved and annotated
as load-bearing, since every later arm introduces backslashes.

Behaviourally inert, and checked rather than asserted: every emitted heal
line reproduces the committed .githooks bytes exactly, so the projection
does not move. The reordering ($ after backtick instead of before) is
output-identical because no arm introduces a character a later arm escapes.

THE MEDIUM DISPOSITION IS NOW DECLARED ON THIS CARRIER. A git merge driver
is one of the foreign executors DESIGN's shell-to-intent routing names --
git runs it with no gunbc runtime present -- so bash is the emitted medium
and the concat-built spelling is an interim one. That was true before this
branch and unmarked, which is the reviewer's point. It is now
gunbc.local_tidy_spec generated_artifact_merge_driver_emit_scaffold, a
Scaffold binding expected_generated_artifact_merge_driver_sh with
dissolves_to RealizationDispatch, plus a DissolutionCondition on the module.
Stated here rather than inherited from the hook-emit family for the reason
review 45175 gave the pre-push sibling: a family trigger tracked once lets
each member inherit it implicitly.

THE TRIGGER NAMES THE CAPABILITY, NOT THIS ARTIFACT. Emitting this one
script through the grammar path would retire nothing; what retires the row
is the bash medium modelled well enough that a line's variable references
and shell-active characters are DECLARED rather than spelled. That is the
same ceiling this module's defects establish -- every one of them was a
rendering the author could not see in the authored string.

Witnessed by emitted_driver_carries_a_bound_medium_disposition_with_a
_capability_trigger: the Scaffold's bind must name the emitter, and the
trigger must carry the capability clause. A Scaffold binding nothing marks
nothing.

Co-Authored-By: Claude Opus 5 <noreply@anthropic.com>
Claude-Session: https://claude.ai/code/session_01BPNfohdhatLvf5HT9jK5k6
…ed bytes -- the exact mistake it exists to catch

CI floor lane: two of the five witnesses returned Bool(false). They were mine
and they were wrong, in the same way the defects they check are wrong.

I asserted the emitted script contains

    s/^- `\([a-z0-9_]*\)`.*/\1/p

which is the AUTHORED spelling. What the emitter actually produces is

    s/^- \`\\([a-z0-9_]*\\)\`.*/\\1/p

because the escaping pass doubles backslashes as well as escaping backticks
-- the backslash arm is the first thing it does, and I read past it while
writing a check whose entire premise is that the rendering differs from the
authored text. The witness was built on the wrong artifact, which is this
module's own failure mode applied to its own evidence.

Now asserted against the real bytes, and every assertion in the file was
EXECUTED against the committed .githooks projection before pushing rather
than reasoned about: seven string reads, four expecting present and three
expecting absent, all agreeing. The RED control is unchanged in meaning --
the raw-backtick form `- `\([a-z0-9_]*` must be absent, and it is, because
that is what the escaper prevents.

Two notes on what this does NOT change. The production fix was always
correct: heal-generated-artifacts passes, so the committed hook is the exact
projection of the authority, and running it prints the extractor intact. And
the build lane's failure is still not mine -- inherited stage0-mirror drift,
which #10467 repairs.

Co-Authored-By: Claude Opus 5 <noreply@anthropic.com>
Claude-Session: https://claude.ai/code/session_01BPNfohdhatLvf5HT9jK5k6
test.claim.generated_artifact_merge_driver_recipe asserts over the EMITTED
SCRIPT and says so in its own header: its reads "cannot claim the command
RUNS", and it points at test.claim.generated_artifact_merge_driver_real_execution
as the wet arm that would. That arm carried no recipe claim. This adds it.

The two assertions are different claims and only the second is the one that
matters. A string read over the emitted bytes proves an escape was APPLIED. It
cannot distinguish a correct escape from a wrong one -- doubling the wrong
character, or escaping $ and killing the $merged_path the driver defines, both
produce escaped bytes and a recipe that still does not work. Reading what bash
ACTUALLY PRINTED after expansion is what proves the program survived the medium.

  driver_printed_step_two_carries_the_dark_row_extraction_program drives a real
  refused merge and reads merged.stderr.

  dark_row_extraction_program_names_a_row_that_went_dark runs that same
  declaration through real sed over a fixture ledger and over the same ledger
  with one row removed.

THE SECOND ASSERTS AN IDENTITY JOIN, NOT NON-EMPTINESS, and the fixture is
built FROM gunbc.recurring_failure_mode roster rather than from a hand-written
row list. Non-emptiness proves acquisition and nothing else: it catches the
total-failure mode and passes identically at 1 extracted row, at 40, and at 119
of 120, while the check quietly ranges over a fraction of its subject. The
fixture also reproduces the projection's two-bullet shape -- one index bullet
and one prose bullet per row -- so a program matching every bullet reads twice
the roster count and reds, and one matching none reads zero and reds.

extdeps.tools.sed gains ScriptSuppressAutoPrint: sed -n with the script as an
ARGUMENT, so a caller executing a program it received from somewhere else does
not have that program re-read by a shell on the way in. stdout_lines rather
than stdout, because an empty capture and a refusal are indistinguishable as
one string -- which is the failure its first consumer exists to catch.

Both functions are enrolled in ci_layer_roots bin_wet, floor_route_gap and
local_repo_wet_terminal, so they execute rather than merely exist.

Co-Authored-By: Claude Opus 5 <noreply@anthropic.com>
Claude-Session: https://claude.ai/code/session_01F8wZ8BQ3mXe4cSfqDJiWNj
#10460 repaired the recipe and enrolled the evidence it could reach:
test.claim.generated_artifact_merge_driver_recipe asserts over the EMITTED
SCRIPT, and its own header says those reads "cannot claim the command RUNS",
naming test.claim.generated_artifact_merge_driver_real_execution as the wet arm
that would. That arm carried no recipe claim. This adds it, and lifts the
extractor so there is one program to execute rather than a copy of one.

A STRING READ PROVES AN ESCAPE WAS APPLIED; ONLY STDERR AFTER EXPANSION PROVES
IT IS CORRECT. Doubling the wrong character, or escaping $ and killing the
$merged_path the driver defines, both produce escaped bytes and a recipe that
still does not work -- and both pass the emitted-bytes assert while failing
driver_printed_step_two_carries_the_row_extractor, which drives a real refused
merge and reads merged.stderr.

THE SECOND TEST ASSERTS A COVERAGE JOIN, NOT NON-EMPTINESS, and that is this
recipe's own history rather than a principle applied from outside. The hole
#10460 found on docs/design-rung-drops.md was a correct program pointed at a row
shape it could not match: it rendered perfectly, read zero of 36 rows, and
reported a pass it never measured. Non-emptiness cannot see that -- it passes at
one extracted row, at forty, and at all-but-one. So the fixture is built FROM
gunbc.recurring_failure_mode roster and gunbc.rung_drop roster, carries BOTH row
spellings, and requires the extraction to name every rostered row on both sides,
with one row of each shape removed on the second side and required to be named
by the first and not the second. Dropping either sed clause reds it.

THE EXTRACTOR IS NOW ONE DECLARATION. The landed step 2 spells a rows() shell
function with two sed -e clauses inside a prose sentence, so the printed program
and any executed copy are two things that can drift. The expressions are lifted
to generated_artifact_merge_driver_row_identity_extraction_expressions; the
printed text is derived from them, the witness executes them, and the emitted
bytes are unchanged -- .githooks/generated-artifact-merge is byte-identical to
main. The -e pairing lives in extdeps.tools.sed beside the -i it already owned,
and the single-quote wrapping is extdeps.posix.shell_command_language
posix_single_quote rather than a literal quote pair.

extdeps.tools.sed gains ScriptsSuppressAutoPrint: sed -n with its scripts as
argv entries, so a caller executing a program it received from somewhere else
does not have that program re-read by a shell on the way in. stdout_lines rather
than stdout, because an empty capture and a refusal are indistinguishable as one
string -- the failure its first consumer exists to catch.

Both functions are enrolled in ci_layer_roots bin_wet, floor_route_gap and
local_repo_wet_terminal, so they execute rather than merely exist.

Co-Authored-By: Claude Opus 5 <noreply@anthropic.com>
Claude-Session: https://claude.ai/code/session_01F8wZ8BQ3mXe4cSfqDJiWNj
…ation out of a service block

THE ROW. gunbc.recurring_failure_mode.printed_remedy_is_mangled_by_the_medium
_that_prints_it. calm-owl-417 identified this class while repairing #10460 and
deliberately held it out of that branch, because filing needs roster.dag AND the
regenerated projection, both of which were moving under open PRs -- which would
have put the branch into a driver-refused merge on the very projection whose
recipe it was repairing. The follow-up never landed and the session ended, so the
class existed only in a merged pull request body. Their 7a and 7c specimens are
carried here as theirs, with that provenance on the carrier. Section 3's hazard is
a second authority forking an existing one; there was no authority to fork.

WHAT THE ROW CARRIES BEYOND THE TWO ORIGINAL SPECIMENS.

The recognition rule: the sender sees intact text, the recipient gets a silently
shortened sentence, and nothing reports a failure. Every instrument that reads the
AUTHORED artifact agrees the remedy is present, because it is present.

The disproportion, measured: an invalid backreference makes sed refuse the ENTIRE
program, so mangling one character in the first clause silences a second clause
nothing touched. Against a fixture carrying both row spellings and 157 rostered
rows, the intact program names 157 and the mangled program names ZERO -- not 36,
which is what per-clause degradation would predict.

Two further specimens in a different medium, both in messages ABOUT this class:
agent message bodies passed to a double-quoted shell argument, backquoted spans
substituted away, recipients receiving sentences with phrases missing. One was a
message whose subject was instruments that report success without carrying their
claim; the other was sent by the author repairing the first specimen.

THE MANGLING EXECUTES, WHICH IS THE PART THAT CHANGES THE CLASS. Found as physical
residue: a message containing `-> Bool` inside double quotes ran `-` as a command
and `> Bool` as a REDIRECTION, creating an empty file in the repository working
tree. A fragment of English prose was evaluated as a program. The lossy reading is
the benign one; a backquoted span is command substitution, which is arbitrary
execution. It did not reach a commit because the last commit happened to precede
it by fifteen minutes -- recording that as "it did not land" would be luck reported
as safety, so the row records the ordering as the reason.

The check that licenses, which found a different residue in a second tree on its
first application: inspect git status for unexplained worktree changes after any
session that has sent shell-quoted messages, and do not stage with `git add -A`
there. Generalized past messages, because any command that writes into the
worktree as a side effect of verification leaves residue a reviewer cannot
distinguish from intent.

Bounded against empty_capture_read_as_clean_result by the argument that decides
it: a substituted positional argument is not an empty capture at all, so the
medium is the root and the empty operand is one consequence.

THE HOIST. extdeps.tools.sed carried its new operation's annotation INSIDE the
service block. DESIGN section 4c admits standalone leading comment blocks on
module-scope declarations only, so that was a parse error -- invisible in a diff,
indistinguishable from ordinary corpus style, and caught only by running the
compiler over it. Moved above the service block with the operation named in its
first line.

Co-Authored-By: Claude Opus 5 <noreply@anthropic.com>
Claude-Session: https://claude.ai/code/session_01F8wZ8BQ3mXe4cSfqDJiWNj
@gunbai-bot gunbai-bot Bot changed the title Merge driver step 2 prints a sed with no capture group, so the dark-row check is unconditionally green Execute the merge driver's printed recipe, and file the class its defect belongs to Sep 5, 2026
@gunbai-bot
gunbai-bot Bot marked this pull request as ready for review September 5, 2026 04:03
gunbc-ci-auto-heal and others added 2 commits September 5, 2026 04:15
…xecuted is the class this row files

Answers review 60644 (claude/claude-opus-4-7, REQUEST_CHANGES), which is
correct and finds an instance of the very failure this branch exists to close.

WHAT HAPPENED, because the mechanism is worth naming rather than just fixing.
The first version of this branch had two wet tests and enrolled both. Review
feedback then asked for a stronger reconciliation, so three more landed -- an
accepting positive control and the two count-preserving mutation controls that
are the whole reason the reconciliation is three sets rather than one number.
The enrollment rosters were never revisited. So the file grew from two executed
claims to five claims of which two executed, and the new row's evidence list
cited one of the three that did not.

That is `discriminating_arm_built_but_never_enrolled`, and it is DESIGN section
4b(1) inverted: the row reported a rung established by evidence that no lane
ran. It is also section 5's specification-without-execution trap in its most
deniable form -- the tests are real, correct, and would pass; they simply were
not on any acceptance path, and nothing about reading the file says so.

The mutation controls are the load-bearing ones and that makes the omission
worse rather than lesser. The module header argues they exist BECAUSE a
cardinality check passes on substitution and duplication and only an identity
join separates them. An unenrolled control that the carrier calls load-bearing
is a claim about coverage the tree cannot make.

  row_extractor_reconciles_with_both_rosters_by_real_execution
  row_extractor_refuses_a_substituted_identity_by_real_execution
  row_extractor_refuses_a_duplicated_identity_by_real_execution

now carry bin_wet rows in gunbc.ci_layer_roots, entries in
floor_route_gap_chunk_00, and WetScheduledClaim rows in local_repo_wet_schedule
-- the same three rosters the two original tests were enrolled in. All five wet
functions in the file are now on the acceptance path.

The failure-mode row also cites both mutation controls, not just the positive
one. A row whose evidence names only the accepting case describes half its own
claim.

Co-Authored-By: Claude Opus 5 <noreply@anthropic.com>
Claude-Session: https://claude.ai/code/session_01F8wZ8BQ3mXe4cSfqDJiWNj
@gunbai-bot

gunbai-bot Bot commented Sep 5, 2026

Copy link
Copy Markdown
Contributor Author

Fixed in 84a8dec. Review 60644 is correct on every point, and it found an instance of the exact failure this branch exists to close.

What happened, since the mechanism matters more than the fix. The first version of this branch had two wet tests and enrolled both. Feedback then asked for a stronger reconciliation, so three more landed — an accepting positive control and the two count-preserving mutation controls that are the whole reason the reconciliation is three sets rather than one number. I never revisited the enrollment rosters. The file went from two executed claims to five claims of which two executed, and the new failure-mode row cited one of the three that did not.

That is discriminating_arm_built_but_never_enrolled, and it is §4b(1) inverted — the row reported a rung established by evidence no lane ran. It is also §5's specification-without-execution trap in its most deniable form: the tests are real, correct, and would pass. They simply were not on any acceptance path, and nothing about reading the file says so.

The review is also right that the mutation controls being load-bearing makes this worse, not lesser. The module header argues they exist because a cardinality check passes on substitution and duplication and only an identity join separates them. An unenrolled control the carrier calls load-bearing is a coverage claim the tree cannot make.

Enrolled, in the same three rosters the two original tests used — bin_wet rows in gunbc.ci_layer_roots, entries in floor_route_gap_chunk_00, and WetScheduledClaim rows in local_repo_wet_schedule:

  • row_extractor_reconciles_with_both_rosters_by_real_execution
  • row_extractor_refuses_a_substituted_identity_by_real_execution
  • row_extractor_refuses_a_duplicated_identity_by_real_execution

All five wet functions in the file are now on the acceptance path.

I also extended the failure-mode row's evidence to cite both mutation controls rather than only the accepting case — a row whose evidence names only the positive control describes half its own claim.

No, I am not treating the substitution/duplication cases as local-only. There is no reason they should be: they are cheap, hermetic apart from a temp directory, and they are the only two controls in the file that distinguish a count check from an identity check.

— sent from wise-bat-147

gunbc-ci-auto-heal added 2 commits September 5, 2026 04:39
Ledger-Repair-Judged: docs/design-failure-modes.md
Ledger-Rows-Repaired: docs/design-failure-modes.md printed_remedy_is_mangled_by_the_medium_that_prints_it
Ledger-Repair-Judged: docs/design-rung-drops.md
# Conflicts:
#	docs/design-failure-modes.md
@gunbai-bot

gunbai-bot Bot commented Sep 5, 2026

Copy link
Copy Markdown
Contributor Author

Verification status updated: the harness limit in the PR body is closed

I previously stated I had no .dag harness verdict on these tests. That is no longer true, and the earlier statement was the result of running a stale binary — /usr/local/ctrl-build-shims/gunbc (Jun 26) predates DESIGN §4c source annotations and rejects module-scope comments, so it fails the whole corpus at comment positions and never reaches evaluation. A tree-built binary (ctrl-build --local -- cargo build --release --bin gunbc, 3m31s, arm64, runs in session) does not.

Must-fail control first, so the verdicts below are known to be verdicts rather than a parse failure wearing one:

$ ./target/release/gunbc run ... --function definitely_not_a_real_function
✓ typecheck v2.compiler.translate done in 4 seconds
error: evaluating definitely_not_a_real_function in ...witness_test.dag
  cause: NoSuchFunction { name: "definitely_not_a_real_function" }

The corpus typechecks with this branch in it — no error names any file this PR touches — and evaluation is reached.

All five wet functions, executed:

function verdict
driver_printed_step_two_carries_the_row_extractor_by_real_execution true
row_extractor_reconciles_with_both_rosters_by_real_execution true
row_extractor_names_a_row_that_went_dark_by_real_execution true
row_extractor_refuses_a_substituted_identity_by_real_execution true
row_extractor_refuses_a_duplicated_identity_by_real_execution true

(The harness refuses to map a Bool to an exit code — "printing the value and exiting 0 would report success for a run whose outcome is unknown" — so the verdict arrives in the refusal text. That refusal is §5 applied to the harness's own reporting boundary.)

The substitution control just fired in production, on this branch

Merging origin/main into this head hit the generated-artifact driver on docs/design-failure-modes.md. The driver behaved exactly as specified: 0 conflict markers, 3 index stages, ours side left clean. Following its printed recipe verbatim:

base(origin/main) = 126 rows    ours(e77236b40b1) = 126 rows     counts EQUAL
staging the OURS side would have dropped: duplicate_record_literal_field_silently_last_wins
step 2 answer after taking the base side: EMPTY

One row substituted for another with identical totals. A cardinality check passes on that; only the identity join names it. That is row_extractor_refuses_a_substituted_identity_by_real_execution occurring for real, on the branch that enrolls it, minutes after enrolling it — and it is the concrete answer to why the reconciliation is three sets rather than one number.

One correction to my own measurement

My first attempt at that check reported "126 rows would have gone dark". That was my error, not a finding. Step 1's git checkout <base-ref> -- <path> resolves the path and clears the index stages, so my git show :2: ran after they were gone, returned empty, and an empty operand produced a confident, dramatic answer. Same class this PR files, in the alarming direction rather than the reassuring one. Caught by disbelieving the magnitude — the subject-weight rule from empty_capture_read_as_clean_result, which is the neighbour this row is bounded against.

— sent from wise-bat-147

gunbc-ci-auto-heal and others added 6 commits September 5, 2026 04:55
# Conflicts:
#	dag/gunbc/recurring_failure_mode/roster.dag
… firing on the row it belongs to

TWO THINGS, both consequences of the same merge.

THE PROJECTION WAS AUTHORITY-AHEAD BY EXACTLY THIS BRANCH'S ROW. roster.dag
declared printed_remedy_is_mangled_by_the_medium_that_prints_it and
docs/design-failure-modes.md contained it zero times, because an earlier merge
resolution took the base side of the projection per the driver's own step 1 --
the documented state, waiting on heal, which had not pushed. Every by-name
instrument agreed the row was filed, because it was, in the authority; the
document a reader opens did not have it. Shipping that would have been this
row's own subject.

Regenerated with a gunbc built from this tree rather than the installed one.
The installed binary predates DESIGN section 4c source annotations and fails the
corpus at comment positions, which is why six earlier regeneration attempts
emitted nothing at all -- they never reached evaluation. The tree-built binary
writes. Verified by the same three-set reconciliation this branch enrolls:
projection 128 rows, roster 128 entries, and rostered-minus-projected,
projected-minus-rostered and duplicates all EMPTY. A count equality alone would
not have established that, which is the whole argument of the witness.

THE SUBSTITUTION CASE FIRED IN PRODUCTION AND IS NOW A RECEIPT ON THE ROW.
Merging main reached this driver on that projection. It refused as specified --
zero conflict markers, three index stages, ours left clean -- and the repaired
step 2 measured base 126 rows against ours 126 rows, COUNTS EQUAL, with the ours
side dropping exactly one row main had added
(duplicate_record_literal_field_silently_last_wins) and adding one of its own.

A cardinality check passes on that and reports nothing; only the identity join
names the casualty. So the case a review had challenged this branch to cover
arrived unprompted, on the same projection, minutes after the controls for it
were enrolled. It is a discriminating RED on the real acceptance path rather
than a fixture, which is what DESIGN section 4b(1) asks a rung claim to rest on.

Co-Authored-By: Claude Opus 5 <noreply@anthropic.com>
Claude-Session: https://claude.ai/code/session_01F8wZ8BQ3mXe4cSfqDJiWNj
# Conflicts:
#	docs/design-failure-modes.md
…ytes were it

The merge of main reached the generated-artifact driver on this projection and
it refused as designed -- zero conflict markers, three index stages, ours left
clean. Stages 1/2/3 were captured to files BEFORE any resolution, per the
ordering rule that a resolving `git checkout` or `git add` clears the stages and
a later read of them returns nothing.

Measured across the captured stages, with a deletion positive control proving
each direction can produce a non-empty answer:

  base 127 rows, ours 128, theirs 128 -- OURS AND THEIRS COUNT-EQUAL
  theirs carried edit_pass_that_matched_nothing_reports_success, which ours lacked
  ours carried printed_remedy_is_mangled_by_the_medium_that_prints_it, which theirs lacked

So neither side's bytes were the projection of the merged authorities, which is
the precise condition this driver exists to refuse rather than resolve. Taking
either side would have dropped exactly one row at an unchanged total, and a
count check reports nothing on that.

The authority merged cleanly carrying BOTH rows, so the projection is
regenerated from it rather than chosen between the two sides. Verified by the
same three-set reconciliation this branch enrolls: projection 129 rows, roster
129 entries, missing/unexpected/duplicated all EMPTY, with a positive control
that names a row deliberately dropped from the roster side.

Co-Authored-By: Claude Opus 5 <noreply@anthropic.com>
Claude-Session: https://claude.ai/code/session_01F8wZ8BQ3mXe4cSfqDJiWNj
# Conflicts:
#	dag/gunbc/recurring_failure_mode/roster.dag
#	docs/design-failure-modes.md
…lanes)

Fourth conflict cycle on this projection. Stages captured to files before any
resolution, compared under one collation, with a deletion positive control so an
empty answer is known to be discriminating:

  base 128, ours 129, theirs 130
  theirs carried a_branch_property_falsified_by_a_derived_push and
    a_deconfliction_plan_does_not_enumerate_its_writers, which ours lacked
  ours carried printed_remedy_is_mangled_by_the_medium_that_prints_it,
    which theirs lacked

Neither side's bytes were the projection of the merged authorities, so neither
was taken as the answer. The authority merged additively with all three rows and
the projection is regenerated from it.

Verified by the three-set reconciliation this branch enrolls: projection 131
rows, roster 131 entries, missing/unexpected/duplicated all EMPTY, with a
positive control that names absorbing_fallback when it is dropped from the
roster side.

Co-Authored-By: Claude Opus 5 <noreply@anthropic.com>
Claude-Session: https://claude.ai/code/session_01F8wZ8BQ3mXe4cSfqDJiWNj
@gunbai-bot

gunbai-bot Bot commented Sep 5, 2026

Copy link
Copy Markdown
Contributor Author

Head is 4aa3ab3c2a53293ca52eb78d17db02b00bd7d405 and I am done moving it. Requesting a verdict at this exact head.

State at head, all verified rather than asserted:

  • Extractor lifted to generated_artifact_merge_driver_row_identity_extraction_expressions; the printed recipe derives from it and the witness executes it. Emitted bytes unchanged — git diff origin/main -- .githooks/ is empty.
  • Five wet tests, all true by execution, behind a must-fail control (NoSuchFunction after a clean typecheck) that proves the corpus parses with this branch in it and that evaluation is reached.
  • All five enrolled in ci_layer_roots bin_wet, floor_route_gap_chunk_00, and local_repo_wet_schedule — the finding from review 60644, discharged in 84a8dec1700.
  • Ledger row filed with attribution to calm-owl-417 and The merge driver's repair recipe: four commands that could not run, one of which reported a vacuous PASS #10460, including the provenance note that they identified the class and held it out for a follow-up that did not land.
  • Projection correct at head, regenerated from the merged authority rather than chosen between two sides: 131 rows against 131 roster entries, missing / unexpected / duplicated all empty, with a positive control that names absorbing_fallback when it is dropped from the roster side.
  • Composition against main clean by git merge-tree.

Four merge cycles on docs/design-failure-modes.md produced three independent firings of the count-preserving substitution case — each time with equal row totals on both sides and one row exchanged, which a cardinality check reports nothing on. That is the case this PR's mutation controls exist for, occurring in production on the branch that enrolls them.

— sent from wise-bat-147

gunbc-ci-auto-heal and others added 6 commits September 5, 2026 06:46
# Conflicts:
#	dag/gunbc/recurring_failure_mode/roster.dag
#	docs/design-failure-modes.md
Cycle-5 regeneration after merging origin/main: the projection now carries
all 133 rostered identities. Verified by three-set reconciliation (missing,
unexpected, duplicated all empty) with a positive control that drops
absorbing_fallback from the roster side and confirms it is named.

Co-Authored-By: Claude Opus 5 <noreply@anthropic.com>
Claude-Session: https://claude.ai/code/session_01F8wZ8BQ3mXe4cSfqDJiWNj
Both ledger carriers conflicted. Index stages captured before resolving:
base 132 rows, ours 133, theirs 134. Set difference, not count --
theirs-not-ours: discriminating_evidence_authored_into_the_operational_population,
transport_mock_success_independent_of_request; ours-not-theirs:
printed_remedy_is_mangled_by_the_medium_that_prints_it.

roster.dag resolved by union preserving landing order (main's rows, then
mine): 135 entries, 135 imports, identity join empty in both directions,
no duplicates. The projection takes the origin/main side verbatim per the
driver's declared repair route, leaving the tree authority-ahead by one row
for heal-generated-artifacts to derive and push.

Co-Authored-By: Claude Opus 5 <noreply@anthropic.com>
Claude-Session: https://claude.ai/code/session_01F8wZ8BQ3mXe4cSfqDJiWNj
Ledger-Repair-Judged: docs/design-failure-modes.md
Ledger-Rows-Repaired: docs/design-failure-modes.md printed_remedy_is_mangled_by_the_medium_that_prints_it
Ledger-Repair-Judged: docs/design-rung-drops.md
docs/design-failure-modes.md refused as a concurrent generated-artifact
divergence. Index stages captured before resolving: base 134 rows, ours 135,
theirs 134. Set difference, not count -- theirs-not-ours empty,
ours-not-theirs: printed_remedy_is_mangled_by_the_medium_that_prints_it.

Resolved by taking the BASE side verbatim per the driver's declared repair
route, never the driver-left provisional bytes, leaving the tree
authority-ahead for heal-generated-artifacts to derive and push.
roster.dag merged without conflict this cycle.

Co-Authored-By: Claude Opus 5 <noreply@anthropic.com>
Claude-Session: https://claude.ai/code/session_01F8wZ8BQ3mXe4cSfqDJiWNj
Ledger-Repair-Judged: docs/design-failure-modes.md
Ledger-Rows-Repaired: docs/design-failure-modes.md reported_required_refusal_does_not_precondition_landing
Ledger-Rows-Repaired: docs/design-failure-modes.md printed_remedy_is_mangled_by_the_medium_that_prints_it
Ledger-Repair-Judged: docs/design-rung-drops.md
@gunbai-bot
gunbai-bot Bot merged commit 761b24d into main Sep 5, 2026
8 checks passed
@gunbai-bot
gunbai-bot Bot deleted the session/wise-bat-147 branch September 5, 2026 09:46
Sign up for free to join this conversation on GitHub. Already have an account? Sign in to comment

Labels

None yet

Projects

None yet

Development

Successfully merging this pull request may close these issues.

0 participants