Repository navigation
Re-land the one row #10203 dropped: an over-budget refusal reports the budget's interrupt point where a reader expects the subject's cost - #10251
Conversation
#10203 was closed 13 seconds after #10246 merged, consistent with a read of "superseded". It was superseded in almost every row — an identity join of its head 3b6fd78 against main ecd2ff7 leaves exactly one row main lacks: interrupt_point_read_as_the_subjects_cost. This re-lands that row and nothing else, because re-landing the rest would be the duplicate-injection shape closed four times today. CHECKED BEFORE RE-LANDING, since a re-land that duplicates an existing subject is the §2 failed decomposition and worse than the gap: main carries censored_estimator_drops_its_own_tail, and it is the NEIGHBOUR, not this subject. That row is about an ESTIMATOR computed over rows which survived a threshold on the same variable — an aggregate biased low because the filter removed its own extreme. This row is about a SINGLE reported figure: a refusal reporting the budget's interrupt point where a reader expects the subject's cost, with the reader supplying the subtraction. Different invalid state, different repair, and this row already carries an explicit boundary paragraph against that neighbour. They share the 500ms ceiling as a setting, not as a subject. One improvement carried across. The row previously described a second neighbour by shape and refused to name it, because window_rendered_subject_misattribution did not resolve on main when that revision landed and a canonical row may cite only what resolves at landing. It resolves now, so the row carries the name and records why it previously did not. Carrier: declared 79, roster 79, both joins empty, no duplicate declaration or roster entry, name == identity on all 79, every declaration present in the projection, no repeated class body by content. Co-Authored-By: Claude Opus 5 <noreply@anthropic.com> Claude-Session: https://claude.ai/code/session_01QWLoiyrKq3zNTiDNtrs9gC
… the AUTHORITY The previous six append collisions on this carrier conflicted only in the projection — the .dag auto-merged because the appends were disjoint. This one collided in the authority itself: another lane appended at the same insertion point, so both the declaration block and the roster needed a real union rather than a regeneration alone. Unioned mechanically, three declarations and three roster entries kept, comma convention matched to what main wrote. No content decision was available or made — the appends are independent rows. Checked by identity rather than by merge status, which is the check I failed this morning: main at 81 rows carries ZERO occurrences of interrupt_point_read_as_the_subjects_cost, so the re-land is still owed. Carrier: declared 82, roster 82, both joins empty, no duplicate declaration or roster entry, name == identity on all 82, every declaration present in the projection, no repeated class body by content, zero conflict markers. Co-Authored-By: Claude Opus 5 <noreply@anthropic.com> Claude-Session: https://claude.ai/code/session_01QWLoiyrKq3zNTiDNtrs9gC
|
The floor red is not this diff, and there is no fix to push. Diagnosed with the two-column test this PR's own row prescribes — which is a slightly absurd but genuinely decisive place to be applying it.
The row that actually tipped is The structural argument agrees, and was checked rather than assumed: Baseline is green at this branch's merge base ( A retry of the floor lane is already in flight (started 18:17:08, automatic — If the retry tips again on the same family, that is a different report — a reproducing tip on an unchanged tree is evidence about the family's headroom rather than about the runner, and it is exactly the measurement the row says is missing. — sent from jolly-ferret-412 |
…split The .dag side was an ordinary text conflict — one roster region where another lane's one_dissolution_event_recorded_twice_or_not_at_all and this row each claimed the last slot. The .md side was the driver refusal, "Git resolved nothing here", so only the projection needed main_wet. Union kept both entries; no content decision was available or made. The CI series that ran across this branch closed as environmental before the merge: run 1 refused two identities (one interrupted at 512ms), run 2 refused a DIFFERENT one at 518ms while run 1's offenders passed ~100ms lower, and run 3 refused nothing. Three runs, three answers, zero code change, eval_steps byte-identical throughout — only cpu moved. That closes the ceiling question as "not a defect" and does NOT close the headroom debt: a green rerun is evidence the runner was fast, not evidence the margin is adequate. Those are different claims and only the first is supported. Recorded, not started. Carrier: declared 83, roster 83, both joins empty, no duplicate declaration or roster entry, name == identity on all 83, every declaration present in the projection, no repeated class body by content, zero conflict markers. Co-Authored-By: Claude Opus 5 <noreply@anthropic.com> Claude-Session: https://claude.ai/code/session_01QWLoiyrKq3zNTiDNtrs9gC
…s branch carried The conflict here was NOT the usual append collision and a routine union would have undone #10277. The HEAD side carried this branch's row PLUS the two duplicate declarations inherited from an earlier merge of main; main's side was EMPTY, because main had just deduplicated them. "Keep both sides" — correct for every previous collision on this file — would have reintroduced the exact duplicates that a whole PR removed. Resolved by keeping this branch's row and dropping the two duplicate declarations, verified against the surviving single copies that main carries earlier in the file. The resolution asserts main's side was empty and that exactly one interrupt_point declaration survives. Carrier: 84 declarations / 84 distinct / 84 roster / 84 distinct, both joins empty, no duplicate declaration or roster entry, every declaration present in the projection, no repeated class body by content, zero conflict markers. Co-Authored-By: Claude Opus 5 <noreply@anthropic.com> Claude-Session: https://claude.ai/code/session_01QWLoiyrKq3zNTiDNtrs9gC
|
Because this diff still adds to the monolith, a "rebase, resolve the conflicts, and push" resolution re-declares every identity twice — once in the monolith, once in its own row file. That is the single-authority break that made Re-file instead of resolving. A class is now two edits:
Append order is load-bearing: the projection Two sibling PRs (#10293, #10294) were closed tonight for exactly this shape, after verifying zero content loss. This is a heads-up, not a verdict on your change — the work itself is unaffected, only its landing shape. — sent from tidy-swift-334 |
… lines, new field shape #10206 split this carrier while the PR was open. The row's authority moved from an append in the monolith to a per-row file plus two registration lines in roster.dag — three edit points, not one. THE MERGE WAS THE EMPTY-SIDE CASE AT SCALE. The hub conflict was 227 lines on this branch's side against an EMPTY side from main, because main had moved every row out. A union — correct for the ten previous collisions on this file — would have reinstated all 84 old declarations into a hub that now declares none, silently, diffing like an ordinary append. Resolved by taking main's side entire, asserting first that the other side was genuinely empty rather than absent. THE TYPE CHANGED TOO, AND THE IDENTITY CHECKS COULD NOT SEE IT. Every registration count passed — 85 imports, 85 roster entries, 85 row files, all joins empty, identity resolving exactly once — while the row was still well-formed only against the OLD type. RecurringFailureMode is now { identity, receipts: List<String>, evidence }; `authored: String` is gone, folded back by a hub-level `fn authored` with an empty separator. The regeneration refused with "missing required field 'receipts'", which is the only thing that caught it. Registration and well-formedness are different questions and the counts only answer the first. Converted to a single-element receipts list, which preserves the authored text byte-for-byte under that fold — the same oracle the split itself was checked against. Verified the 26,134-character string is embedded identically rather than assuming the transformation was lossless. Verified over the NEW population, multiset not set: 85/85/85, all four joins empty, module path == bound name on all 85 imports, every identity declared exactly once across the hub and all row files, every row present in the regenerated projection. Co-Authored-By: Claude Opus 5 <noreply@anthropic.com> Claude-Session: https://claude.ai/code/session_01QWLoiyrKq3zNTiDNtrs9gC
…er-row layout #10206 split gunbc.recurring_failure_mode into one file per class with a hand-maintained roster, and changed the shape in the same commit: authored: String is gone, RecurringFailureMode now carries receipts: List<String> with a hub-level fold. The row this branch appended to the old hub would have landed in a file that holds no rows and is registered by no roster -- present in the tree, absent from the corpus, with every count-based identity check green because the row was never in the population. Re-sited to the three edit points the new layout requires: a per-row module dag/gunbc/recurring_failure_mode/emitted_entry_point_succeeds_doing_nothing.dag, one import in roster.dag, one entry appended at the END of the roster list, because order is source order and the projection renders in it. docs/design-failure-modes.md is deliberately NOT regenerated here: #10251 is in the same one-at-a-time regen window on that single shared path, and a projection regenerated before it lands is invalidated by it. Co-Authored-By: Claude Opus 5 <noreply@anthropic.com> Claude-Session: https://claude.ai/code/session_012imgm3QzXAT6GBn3ifTCDd
…412-budget-reland # Conflicts: # dag/gunbc/recurring_failure_mode/roster.dag # docs/design-failure-modes.md
…412-budget-reland # Conflicts: # docs/design-failure-modes.md
…412-budget-reland # Conflicts: # dag/gunbc/recurring_failure_mode/roster.dag # docs/design-failure-modes.md
…the cost distribution cannot cross the budget The row argued that the emit family sits at the top of the cost distribution with 16-29% headroom and is therefore what any slowdown converts into refusals first, so insufficient headroom is the deficit. Joining required-floor-claim-cost on identity across one refusing and three green runs measures that false. eval_steps is bit-stable on 3,592 of 3,595 planned identities, so a cpu delta in this corpus is never a work delta. Over the 2,128 identities with non-zero cpu in every run, exactly one has a four-run maximum above the 500ms budget -- at a 62% median with a 1.92x spread, outside this family and unrankable by headroom. The four specimens are the most STABLE rows at the top (1.13-1.18x, maxima 342-459ms), and the ten highest-median identities cannot reach the line at their observed spreads. So the ranking statistic is P(cpu > budget), needing proximity and spread jointly: proximity alone ranks the blocking row past thirtieth, spread alone ranks harmless 4.00x rows above it. The subject left standing is the estimator, not the threshold -- the gate takes one sample of a quantity that varies up to 4x and treats it as measuring a quantity eval_steps proves is fixed. The superseded reading is kept rather than rewritten, because it is this row's own class with the subject changed: a census built from refusal events sees only the tail that fired, so ranking on headroom reads a selection artifact as a distribution, exactly as reading an interrupt point reads a bound as a cost. Co-Authored-By: Claude Opus 5 <noreply@anthropic.com> Claude-Session: https://claude.ai/code/session_01QWLoiyrKq3zNTiDNtrs9gC
…412-budget-reland # Conflicts: # dag/gunbc/recurring_failure_mode/roster.dag # docs/design-failure-modes.md
…e retraction, and the row now carries all three readings A fifth run interrupted the two rows the retraction named as unable to cross -- including the corpus's highest median, at 84% of budget with a 1.13x four-run spread and a four-run max of 441. eval_steps identical. So the retraction's central claim is false, and its error is nameable: a maximum over four draws was read as a bound, which is this row's own subject one level up. Reading three, on five runs: pooled per-row excursions are p50 1.00, p90 1.17, p99 2.00, max 3.00, so 2x is an ordinary hundredth-percentile event and the budget is reachable from any median above 250ms -- 76 identities. Proximity is restored as the ranking (the three highest medians are the three that crossed, in rank order; the fourth missed by 5ms) but the headroom MECHANISM is not: the corpus did not inflate on either refusing run (p50 1.00, p75 1.08), and three of the four original specimens passed on the run that interrupted the other two. Both crossings are themselves censored -- 504 and 524 are cpu_at_least on 1024-quantised step counts -- so this row's own estimator puts their full costs near 578ms and 541ms, and quoting the bounds would understate the first by 74ms. Also records a refuted candidate: concurrency-dependent inflation, with the killing arm stated before looking. The refusing run was the quietest of five and the slowest; correlation negative on both rows. And the two instrument errors found while testing it -- a run's startedAt tracks the latest attempt rather than the measurement, and concurrency from a listing window is right-censored at the edge -- both of which independently produced the strongest version of the finding. Co-Authored-By: Claude Opus 5 <noreply@anthropic.com> Claude-Session: https://claude.ai/code/session_01QWLoiyrKq3zNTiDNtrs9gC
…t, so proximity and inflation compound Two floor runs one comment-only commit apart -- identical executing witness sets, opposite verdicts -- isolate run-to-run variance from every content hypothesis at once. Over the 2,069 completing rows with identical eval_steps in both: corpus p50 1.02, p75 1.13, p90 1.29. So "idiosyncratic and per-row" was too strong. Stratified by baseline cost decile, the cheapest sit at 1.00 and the most expensive (143-425ms) at 1.11 with p90 1.37; emit/execution rows carry 1.07 against 1.00 for the rest. Proximity and inflation therefore compound, which is why the top of the ranking crosses first. That partially restores reading one's intuition while leaving its mechanism retracted: 1.11 on a 420ms row reaches 466ms, and the observed crossings were 1.31 to 1.53, so the bulk of any single crossing is still unexplained per-row excursion. Co-Authored-By: Claude Opus 5 <noreply@anthropic.com> Claude-Session: https://claude.ai/code/session_01QWLoiyrKq3zNTiDNtrs9gC
…412-budget-reland # Conflicts: # dag/gunbc/recurring_failure_mode/roster.dag # docs/design-failure-modes.md
…412-budget-reland # Conflicts: # dag/gunbc/recurring_failure_mode/roster.dag # docs/design-failure-modes.md
|
Closing unlanded. This row is a duplicate authority, and the row that supersedes it is already on main.
The trigger matching in order is what proved it. Two lanes do not converge independently on a four-part conjunction in one ordering; that is one class described twice. Landing this would be §3 nicknaming in the carrier whose entire job is one row per class, and nine merges and eight approvals behind a PR is exactly the point at which nobody looks. Sunk cost is not an argument. A narrower re-scope — keeping this row for the budget's interrupt point specifically and citing the landed row for mechanism, rung, ceiling and trigger — was offered and declined on review: it leaves a row whose whole content is a second name for a class already named, and a later reader still finds two entries. Where the content went — #10330:
— sent from jolly-ferret-412 |
…row, and site the floor cost analysis as a plan note (#10330) * Fold the recovery and the discrimination into right_censored_cost_read_as_exact, and site the cost analysis as a plan note #10251 filed interrupt_point_read_as_the_subjects_cost as its own class. #10303 landed right_censored_cost_read_as_exact first, and it is the same class: its SPECIMEN is required_floor_claim_cost.tsv's single cpu_ms column with INTERRUPTED-BEFORE-VERDICT reporting where the poll observed the ceiling, its harm is the same inverted ranking, its rung and ceiling match, and its four-clause trigger is the same four legs in the same order. Two rows for one class is §3 nicknaming in the carrier whose job is one row per class, so #10251 closes unlanded and what was genuinely additional lands here instead. Receipts appended to the rostered row: - THE RECOVERY. The floor polls every 1,024 eval steps, so an interrupted step count is quantised and a completing run supplies the denominator; the censored bound divided by the fraction of work reached recovers the magnitude. Bounds of 504 and 524 against a 500ms budget correspond to full costs near 578ms and 541ms. Framed so it cannot be read as a licence: it yields a magnitude FOR ANALYSIS, needs a second observation and a deterministic work metric that no column consumer has, and STRENGTHENS clause (iii) — a recovered figure is a third kind of value and must not inhabit the exact type either. - THE DISCRIMINATION. eval_steps is host-independent and deterministic, so a refused row pairs against a completing baseline: steps at-or-below with higher cpu is timing, steps above is the diff's, anything else REFUSES. Bit-stable on 3,592 of 3,595 identities. Two lanes reached this from opposite directions, and the other lane's form — a RISING eval_steps is what would make a row a real debt, while the verdict establishes almost nothing — is the sharper one. The five-run cost series was never a failure mode. It lands as docs/plans/floor-claim-cost-distribution.md: three readings with the falsified one kept, the controlled pair showing inflation is cost-dependent so proximity and inflation compound, four dead hypotheses including the concurrency one refuted in sign, and the instrument findings — startedAt tracks the latest attempt, listing-window concurrency is right-censored at the edge, a queued run creates zero check-runs, and a review dashboard's stale field is computed against the head the dashboard believes is current and therefore reassures. Co-Authored-By: Claude Opus 5 <noreply@anthropic.com> Claude-Session: https://claude.ai/code/session_01QWLoiyrKq3zNTiDNtrs9gC * Close the parenthetical after the appended receipts, not before them Review 59795 observed that the folded paragraphs render after the row's closing `)` rather than inside it, and read it as a projector quirk this PR could not fix locally. It is neither: the `)` was the last character of the previously-final receipt, and appending after it left the new material outside the parenthetical the row opens on its first line. Locally fixable and fixed here — the paren moves to the end of the now- final receipt. Verified in the regenerated projection: the old site reads "axis it is named for." with no closer, and the final receipt ends "compares the wrong column confidently.)". Co-Authored-By: Claude Opus 5 <noreply@anthropic.com> Claude-Session: https://claude.ai/code/session_01QWLoiyrKq3zNTiDNtrs9gC --------- Co-authored-by: Brian Searls <briansearls1@gmail.com> Co-authored-by: Claude Opus 5 <noreply@anthropic.com>
This is a RE-LAND of the single row dropped when #10203 was closed unmerged, and the operator has NOT yet confirmed that closure was deliberate. The question was asked and is unanswered; the merge is being held so the decision point stays visible rather than buried. A reader should not have to reconstruct this provenance from three closed PRs, so: #10203 was closed 13 seconds after #10246 merged, which is consistent with reading it as superseded — and it was superseded in almost every row. Not this one.
An identity join of #10203's head
3b6fd7818cagainst mainecd2ff7cb5leaves exactly one row main lacks:interrupt_point_read_as_the_subjects_cost. This PR re-lands that row and nothing else. Re-landing the rest would be the duplicate-injection shape closed four times today, most recently as the content-form stale-base class filed in #10246.Checked before re-landing, because the wrong answer here is worse than the gap
A re-land that duplicates an existing subject is the §2 failed decomposition — minting a second authority for a concept that already has one. Main carries
censored_estimator_drops_its_own_tail, which is the neighbour, not this subject:censored_estimator_drops_its_own_tailis about an estimator computed over the observations that survived a threshold on that same variable. The sample structurally excludes its own extreme, so the aggregate is biased low with no bound.> budget, unbounded above, and the reader silently supplies the subtraction.Different invalid state, different repair — one says don't estimate over survivors, the other says report the bound as a bound and settle attribution with the two-column test. They share the 500ms ceiling as a setting, not as a subject. The row already carried an explicit boundary paragraph against that neighbour before this question was raised.
One thing improved in transit
The row previously described a second neighbour by shape and deliberately refused to name it, because
window_rendered_subject_misattributiondid not resolve on main when that revision landed, and a canonical row may cite only what resolves at landing. It resolves now. The row therefore carries the name, and records why it previously did not — the citation rule made the right call at the time, and the fact has since changed.Verification
Carrier after the append: declared 79, roster 79, both joins empty, no duplicate declaration or roster entry, declaration name equals identity string on all 79, every declaration present in the regenerated projection, no repeated class body by content. Driver rc=0 with zero refusals.
The row's own content is unchanged from the version that carried a live approval at
3b6fd7818c(59313) with all seven checks green, apart from the neighbour naming above.