Skip to content

docs(translation): #202 EN post-release campaign plan — planning-only, 0 CSV write - #809

Merged
jsboige merged 1 commit into
masterfrom
docs/202-en-campaign-plan
Jul 15, 2026
Merged

docs(translation): #202 EN post-release campaign plan — planning-only, 0 CSV write#809
jsboige merged 1 commit into
masterfrom
docs/202-en-campaign-plan

Conversation

@jsboige

@jsboige jsboige commented Jul 15, 2026

Copy link
Copy Markdown
Contributor

What

Planning-only doc structuring the post-v0.9.0 EN translation campaign (epic #202, dispatch y20p0t [PRIMAIRE]).

DoD met: docs-only, 0 CSV write, 0 code change. The campaign itself stays deferred post-release (arbitrage jsboige) — this PR only structures it.

Why

#795 proved the EN core prose (text/desc/example) is 100% clean — there is nothing to gain from a bulk re-translation (risk of regression only). But two Fallacies fields are materially under-filled and ARE used in active templates, so they have real rendered-card impact:

Field Filled Empty Used in
Simple_name_en 61/1408 (4%) 1347 Memo + Fallacies Face
political_example_en 35/1408 (2%) 1373 Fallacies Face (political variant)

All other EN datasets are at/near 100%.

Plan summary (full detail in the doc)

Verification

  • Inventory measured live from prod CSVs at master cf8cb0d8 (2026-07-15).
  • Template usage confirmed via grep (Cards/Fallacies/Argumentum_Fallacies_Face_*.json, Cards/Memo/Argumentum_Memo_*.json).
  • No code change → no build/test impact (docs-only).

— po-2024 (dispatch y20p0t, planning-only GO)

…, 0 CSV write

Structures the post-v0.9.0 EN translation campaign as a plan doc (DoD:
docs-only, no CSV writes). The campaign itself stays deferred post-release
(arbitrage jsboige).

Empirical inventory computed live from the prod CSVs at master cf8cb0d:
- Fallacies Simple_name_en: 61/1408 filled (1347 empty, 96% margin)
- Fallacies political_example_en: 35/1408 filled (1373 empty, 98% margin)
- Both fields ARE used in active templates (Cards/Fallacies/Argumentum_Fallacies_Face_*.json,
  Cards/Memo/Argumentum_Memo_*.json — confirmed via grep) → real rendered-card impact.
- All other EN datasets at/near 100% (#795 proved core prose is clean).

The doc designs:
- gpt-5.5 prompt per field (/v1/responses, reasoning.effort=low, empty-only,
  cell-by-cell verified) reusing the existing DatasetUpdater scaffold
  ("Translate Fallacies to English empty-only 0-shot" task pattern).
- 5 gated tranches (T1/T2 Simple_name_en pilot+bulk, T3/T4 political_example_en
  pilot+bulk, T5 trivial Virtues hierarchy residual), ~2725 calls total.
- Post-tag execution order; explicit out-of-scope (no bulk re-translation,
  no link_* = #192, no #804 Phase 4 regen, #415 history-rewrite INTERDICTED).

Dispatch y20p0t [PRIMAIRE] #202-prep.

Co-Authored-By: Claude-Code <noreply@anthropic.com>
@jsboige
jsboige merged commit 940acbe into master Jul 15, 2026
3 checks passed
@jsboige
jsboige deleted the docs/202-en-campaign-plan branch July 15, 2026 17:08
jsboige added a commit that referenced this pull request Jul 15, 2026
…teria (docs-only) (#810)

Companion to 202-en-campaign-plan.md (#809). Defines the byte-exact
spot-check criteria ai-01 applies per pilot tranche before a bulk GO,
so campaign execution is turnkey once the register decision (A/B) lands.

6 checks per 20-cell pilot:
1. Empty-only invariant (byte-exact, HARD — 0 clobber)
2. Meaning fidelity vs text_fr/desc_fr (semantic, sample >=5)
3. Register/style (field-specific: Simple_name <=5 words; political_example
   neutral/non-defamatory, matches ratified A/B)
4. MT-garbage sweep (3-dim, per-token counts — memo false-zero)
5. Language purity (0 FR/RU leakage in new outputs)
6. Encoding/round-trip (CRLF+BOM, no quoting drift — memo byte-exact-insertion)

GO/REVISE/NO-GO thresholds; hard gates = clobber + encoding drift.
Post-bulk regression gate: re-run #795 drift audit + byte-check ALL lang
columns (zh #761 lesson) + rendered-source freshness (clobber harvests).

#202 execution stays deferred post-v0.9.0 (jsboige arbitration). Docs-only,
0 CSV write, 0 code change.

Dispatch mf4nyb [SECONDAIRE].

Co-authored-by: Your <your.email@example.com>
Co-authored-by: Claude-Code <noreply@anthropic.com>
Sign up for free to join this conversation on GitHub. Already have an account? Sign in to comment

Labels

None yet

Projects

None yet

Development

Successfully merging this pull request may close these issues.

1 participant