docs(i18n): taxonomy cross-lang terminological coherence scan (release sign-off, 0 finding) - #679
Merged
Merged
Conversation
…e sign-off, 0 finding) Secondaire of dispatch `lofjtd` (ai-01). Read-only coherence scan of Fallacies taxonomy family/subfamily labels across 7 non-FR languages - the intra-lang consistency axis not covered by scanner #647 (FR-contamination) nor #192 (FR-relative terminology). Result: 0 inconsistencies at every meaningful granularity: - Family_<lang> by (Famille, Sous-Famille): 29 groups, 0 inconsistent - Subfamily_<lang>: 21 groups, 0 inconsistent - Subsubfamily_<lang>: 63 groups, 0 inconsistent - FR Sous-Famille -> consistent Subfamily_<lang>: 21, 0 inconsistent The apparent Famille-level divergence (7/8 families) is intentional sub-family-level localization, not a data-quality defect. Complements #667 (scanner #647 = 0) and #192 on the three i18n axes. Release Fallacies taxonomy surface: coherent + uncontaminated. 0 CSV write, 0 gpt-5.5 call. Read-only. Co-Authored-By: Claude-Code <noreply@anthropic.com>
This file contains hidden or bidirectional Unicode text that may be interpreted or compiled differently than what appears below. To review, open the file in an editor that reveals hidden Unicode characters.
Learn more about bidirectional Unicode characters
Sign up for free
to join this conversation on GitHub.
Already have an account?
Sign in to comment
Add this suggestion to a batch that can be applied as a single commit.This suggestion is invalid because no changes were made to the code.Suggestions cannot be applied while the pull request is closed.Suggestions cannot be applied while viewing a subset of changes.Only one suggestion per line can be applied in a batch.Add this suggestion to a batch that can be applied as a single commit.Applying suggestions on deleted lines is not supported.You must change the existing code in this line in order to create a valid suggestion.Outdated suggestions cannot be applied.This suggestion has been applied or marked resolved.Suggestions cannot be applied from pending reviews.Suggestions cannot be applied on multi-line comments.Suggestions cannot be applied while the pull request is queued to merge.Suggestion cannot be applied right now. Please check back later.
Summary
Secondaire of dispatch
lofjtd(ai-01). Cross-language terminological coherence scan of the Fallacies taxonomy — the intra-lang consistency axis not covered by scanner #647 (FR-contamination) nor #192 (FR-relative terminology). Read-only, 0 finding, release sign-off. 0 CSV write, 0 gpt-5.5 call.The gap not yet covered
#647 catches "EN cell still contains FR text". It does NOT catch "EN says 'Fallacy' here but 'Sophism' in another row with the same FR
Famille". This scan checks that.Method + result (code=truth)
Group by FR source key → set of localized renderings per lang → flag any group where a lang has >1 value:
Family_<lang>by (Famille, Sous-Famille)Subfamily_<lang>Subsubfamily_<lang>Sous-Famille→ consistentSubfamily_<lang>Verdict: every FR source label maps to exactly one localized rendering per language, across 1408 rows × 7 langs (EN/RU/PT/ES/AR/FA/ZH). The naïve Famille-level divergence (7/8 families, e.g. FR
Tricherie→ ZH人性偏见/作弊/偏见思维) is intentional sub-family-level localization — the localized tree branches differently than FR's umbrella — not a defect. The tight (Famille, Sous-Famille) check confirms this.Release implication
Fallacies taxonomy i18n surface = coherent (this scan) + uncontaminated (#667). Both 0 finding. 8-language bundle ships with consistent family/subfamily labeling.
Scope (not findings)
link_*URL gap → docs(i18n): #192 link_* coverage research track (own the URL gap) #600/docs(taxonomy): #600 link census reproducibility — independent re-run byte-identical #606, human-research axis, not terminological.Reproducible
build_coherence_scan.py(scratchpad) — read-only grouping scan, deterministic. 0 write.Relates #141, #192, #667, #600, #606. Base
d5913862.🤖 Worker po-2024 (dispatch
lofjld, secondaire)