Skip to content

docs(taxonomy): #847 AIF audit Layer (c) reference-reachability (staged, read-only) — MEASURE not verdict - #853

Merged
jsboige merged 1 commit into
masterfrom
docs/847-layer-c-reference-reachability
Jul 23, 2026
Merged

docs(taxonomy): #847 AIF audit Layer (c) reference-reachability (staged, read-only) — MEASURE not verdict#853
jsboige merged 1 commit into
masterfrom
docs/847-layer-c-reference-reachability

Conversation

@jsboige

@jsboige jsboige commented Jul 22, 2026

Copy link
Copy Markdown
Contributor

What

#847 AIF structural audit — Layer (c) reference-reachability acompte (dispatched by ai-01 msg-gvpl31, tick 87). This is the input the layers-(a)+(b) acompte (#850) deferred — the per-node evidential signal that turns each A-candidate / c-DEFER tag from the earlier measure into a bridge-vs-defect reading (method doc §2.2, §4, §6). MEASURE, not verdict. INPUT for ai-01 synthesis, GATED jsboige ratification. 0 prod-CSV. Companion to 847-acompte-homogeneity.md (#850).

Instrument (operational, code=truth)

Classifies each of the 2,579 link_<lang> reference cells (8 langs × Fallacies taxonomy) into one bucket:

Bucket How Reading
wikipedia-stable deterministic allowlist (WMF: wikipedia/wikisource/wiktionary/wikiquote) trustworthy skeleton
archive-repointed deterministic (web.archive.org / archive.is / archive.today) rescued — alive by construction
dictionary-known deterministic allowlist (logicallyfallacious, skepdic, fallacyfiles, don-lindsay-archive, ditext, …) named decaying class
alive / dead / unknown read-only HEAD probe (curl, polite UA, 8s timeout), 2xx/3xx = alive, 4xx/5xx = dead, else unknown long-tail evidential risk
unparseable malformed URL

Conservative by design: deterministic allowlist covers ~93% of refs; the live HEAD probe runs only on the 186 long-tail refs (the evidential-risk class). unknown is never guessed; dead is stated as an upper bound (HEAD-noise caveat, §6).

Headline (2,579 refs)

Bucket n %
wikipedia-stable 2,127 82.5%
dictionary-known 225 8.7%
dead (long-tail HEAD) 134 5.2%
alive (long-tail HEAD) 51 2.0%
archive-repointed 39 1.5%
unknown / unparseable 3 0.1%

Reading. The taxonomy is overwhelmingly Wikipedia-anchored (82.5%). The method doc's §2.2 worry ("dictionary links decaying, several already dead") is confirmed and quantified: 134 dead long-tail refs + a named dictionary class (8.7%) decaying in place.

Per-node findings (MEASURE, not verdict)

  • Family: Influence + Tricherie = Wikipedia skeleton (89-91% wiki, low entropy); Insuffisance = highest evidential risk (63% wiki, 17% long-tail, H=0.68); Erreur de raisonnement = only family at 25% dictionary.
  • Sub-family: Comparaison fallacieuse (44% dict), Mauvaise composition (32%), Mauvaise déduction (28%), Changement de cap (31% long-tail).
  • Sub-sub (leaves, sharpest): Comparaison abusive (80% dictionary), Définition inconsistance (50% long-tail), Argument d'autorité (47% long-tail).

Cross-layer value (§5 — the point of layer c)

Layer (c) is read against #850's layers (a)+(b), never alone:

#850 node Layer-(c) here Synthesis input
Mauvaise déduction (a-high/b-low) dictionary-heavy 28% genuine tree-tension AND evidentially exposed → real arbitration
Ad hominem (c-DEFER in #850) Obstruction 69% wiki, well-anchored stays bridge-node — not explained by evidential decay
Influence (only coherent family) 89% wiki skeleton coherent and best-anchored → high confidence

This converts the acompte's c-DEFER tags from uncheckable placeholders into a per-node evidential profile.

Caveats (honest, §6)

  1. HEAD-noise: some hosts 4xx/5xx on HEAD but 200 on GET (e.g. rationalwiki returned 503 on HEAD). These register dead/unknown → 5.2% dead is an upper bound. GET refinement deferred (heavier, not read-only-HEAD).
  2. Allowlist opinionated; misclassification within the same evidential class is cosmetic.
  3. No GET-body validation (domain-squatted 200 still reads alive) — full evidential audit is human work, out of scope.
  4. Label-distortion flag (method §2.1) not computed here — separate first-pass, mechanical-flag-only, never applied.

Governance

  • MEASURE, not verdict. 0 reorganisation wording. Synthesis + verdict = ai-01.
  • 0 prod-CSV write (T&A freeze). Docs-only artefacts.
  • Read-only HEAD probes (no write/mutation). DatasetUpdater untouched.

Artefacts (docs-only)

  1. docs/taxonomy/847-layerc-reference-reachability.md — narrative + tables
  2. docs/taxonomy/847-layerc-reference-reachability.csv — per-node rollup (90 rows)
  3. docs/taxonomy/847-layerc-reference-reachability-refs.csv — per-reference (2,581 rows)

Refs

🤖 Generated with Claude Code

…ed, read-only) — MEASURE not verdict

Layer-(c) instrumentation dispatched by ai-01 (msg-gvpl31, tick 87). This is the
input the layers-(a)+(b) acompte (#850) deferred — the signal that turns each
A-candidate / c-DEFER tag into a bridge-vs-defect reading (method doc §2.2, §4).

Instrument: classifies each of the 2,579 link_<lang> reference cells (8 langs x
Fallacies taxonomy) into {wikipedia-stable, archive-repointed, dictionary-known,
alive, dead, unknown, unparseable}. Deterministic allowlist for WMF/archive/
named-fallacy-dictionary hosts (~93% of refs); read-only HEAD probe (curl, polite
UA, 8s timeout) only on the 186 long-tail refs (evidential-risk class). Unknown is
NEVER guessed (conservative); dead is an upper bound (HEAD-noise caveat).

Headline: wikipedia-stable 82.5% (trustworthy skeleton) + dictionary-known 8.7% +
dead long-tail 5.2% (134 refs — direct quantification of the method doc "dictionary
decay") + archive-repointed 1.5% + alive long-tail 2.0%.

Per-node rollup surfaces the shape of evidential risk: Influence/Tricherie are the
Wikipedia skeleton (89-91% wiki, low entropy); Insuffisance is highest-risk (63%
wiki, 17% long-tail, H=0.68). Sub-sub leaves concentrate the risk (Comparaison
abusive 80% dictionary; Définition inconsistance 50% long-tail).

Cross-layer value (§5): Mauvaise déduction — genuine tree-tension in #850
(a-high/b-low) AND dictionary-heavy (28%) here = real arbitration, not tradition
artifact. Ad hominem (c-DEFER in #850) sits on a reasonably-anchored Obstruction
base (69% wiki) = stays bridge-node. This is what converts c-DEFER placeholders
into a readable per-node evidential profile.

Governance: MEASURE not verdict. 0 prod CSV. 0 reorganisation wording. Synthesis +
verdict = ai-01. DatasetUpdater untouched. Label-distortion flag (method §2.1)
deferred — separate first-pass, never applied.

Artefacts (docs-only):
- docs/taxonomy/847-layerc-reference-reachability.md (narrative + tables)
- docs/taxonomy/847-layerc-reference-reachability.csv (per-node rollup, 90 rows)
- docs/taxonomy/847-layerc-reference-reachability-refs.csv (per-reference, 2581 rows)

Relates #847 #498 #850. Base master f70b20d.

Co-Authored-By: Claude-Code <noreply@anthropic.com>
@jsboige
jsboige merged commit 9f3caae into master Jul 23, 2026
3 checks passed
@jsboige
jsboige deleted the docs/847-layer-c-reference-reachability branch July 23, 2026 02:45
jsboige added a commit that referenced this pull request Jul 23, 2026
…d, read-only) — MEASURE not verdict (#857)

Layer-(c) second installment dispatched by ai-01 (msg-20260723T024907-2l19wg, tick
89). Completes the §2.1 half (internal-node labelling constraint); companion to #853
(§2.2 reference-reachability, merged 9f3caae). Read-only, 0 run live, 0 prod CSV.

§2.1 has two bullets; only one is mechanically computable:
- (2) lexical availability (borrowed_from_lang): YES — grounded in the taxonomy's
  own Latin column + morphology. Finds Ad hominem (raw Latin) among the 39 grouping
  nodes. Axis-C lexical pressure on the grouping layer = LOW (1/39 raw-Latin).
- (1) distorted scope (stretched/narrowed): NO — semantic. Two proxies tested &
  non-discriminatory (documented, NOT used): concept-dispersion SATURATED (median
  0.94-1.00, each fallacy is its own Wikipedia article); EN-grouping 1:1 everywhere
  (EN columns are a FR translation, not an independent tradition tree -> no axis-C
  structural signal in the grouping layer).

label_fit: 1 borrowed (Ad hominem) + 38 unflagged (no mechanical flag, NOT a
verified scope match). Cross-ref acompte -> combined_reading: Ad hominem = borrowed
x A-candidate = PRIORITY arbitration candidate -> tradition-divergence (bridge-node,
do NOT bill as defect), corroborates #853 (Obstruction base 69% wiki-anchored).
Resolves the c-DEFER on Ad hominem. The other 6 A-candidates carry no mechanical
label signal -> genuine axis-A or human semantic read; layer (c) bills none of them
as tradition artifacts.

Governance: MEASURE not verdict. 0 prod CSV. 0 reorganisation wording. borrowed is
lexical fact (conservative); unflagged = absence of flag, not a positive verdict.
Synthesis + verdict = ai-01/jsboige.

Artefact: docs/taxonomy/847-layerc-label-distortion.{md,csv} (39 nodes).

Relates #847 #498 #853 #850. Base master 82a1e02.

Co-authored-by: Your <your.email@example.com>
Co-authored-by: Claude-Code <noreply@anthropic.com>
jsboige added a commit that referenced this pull request Jul 25, 2026
…ai-01) (#904)

Closes the three-installment measurement chain (#850 layers a+b, #853 layer c
reference-reachability, #857 layer c label-distortion) with the synthesis those
docs explicitly deferred to ai-01.

Verdict: the tree survives the audit. NO restructuring proposed.
- Layer (b) dominates (65.5% fail-loud): most heterogeneity is MATERIAL
  resistance, not authorial grouping error. "Changes at the margin" quantified.
- Ad hominem = cross-cultural BRIDGE node (sole `borrowed` of 39). c-DEFER
  lifted; annotate the seam, do NOT flatten. Ratifies #845's deferred decision.
- Axis-C lexical pressure on the grouping layer is LOW (1/39) and the EN
  grouping tree is a 1:1 FR calque -> tradition divergence lives at the LEAF
  reference level, not in the groupings.
- Applies a minimum-n confidence gate the acompte's n>=3 threshold lacks:
  6 axis-A candidates reduce to 2 for jsboige (Biais naturels, Pensee biaisee).
  Argument bacle is EVIDENCE-BLOCKED (46% wiki / 22% longtail) -> repair refs
  before arbitrating. Erreur de raisonnement watch-only; 2 more thin at n=4-5.
- The chantier works, measurably: folding native-rich clusters moves fail-loud
  -7.4 pp and flips Obstruction + Erreur mathematique from B+A to A-candidate.
  Next clusters: Abus de langage, Tricherie, Insuffisance.
- Reference-repair backlog (S5) is actionable WITHOUT jsboige.

Records the honest caveat that `unflagged` is the absence of a flag, never a
positive scope match: the stretched/narrowed half of the axis-C constraint is
unmeasured and mechanically unmeasurable (both proxies failed, documented #857).

VERDICT, not application. 0 prod-CSV write. All structural change stays gated
on jsboige ratification.

Refs #847 #498 #845 #846 #850 #853 #857

Co-authored-by: Claude Opus 5 (1M context) <noreply@anthropic.com>
Sign up for free to join this conversation on GitHub. Already have an account? Sign in to comment

Labels

None yet

Projects

None yet

Development

Successfully merging this pull request may close these issues.

1 participant