Skip to content

docs: add the accelerated evolution skills post and retitle the bias post - #478

Merged
allxsmith merged 6 commits into
mainfrom
docs/380-ai-loop-keeps-skills-honest
Aug 7, 2026
Merged

allxsmith merged 6 commits into
mainfrom
docs/380-ai-loop-keeps-skills-honest

Conversation

@allxsmith

@allxsmith allxsmith commented Aug 5, 2026 •

Copy link
Copy Markdown
Owner

Summary

  • New post "Accelerated Evolution for Agent Skills" (/blog/ai-loop-keeps-skills-honest), written as one self-contained story: the training-bias problem retold for a stranger, the seven shipped skills, graded cold-start exam runs, ten generations of accelerated evolution (85/100 baseline → 95.2 mean while cost fell 43%), replayable agent runs in Storybook, the review + CI guards that hold the line between generations, and the loop's still-open findings ([Bug] bulma-ui: color props typecheck values with no shipped CSS; Box color falls through to has-text-* #367–[Feature] create-bestax: scaffold polish — gitignore *.tsbuildinfo, PM-neutral next-steps, strictPort fallback #371) plus its stated limits.
  • Cover plus one hero per section (7 SVG+PNG pairs) authored against docs/scripts/pixel-cover-lib.mjs and rasterized through rasterize:cover; the cover headline reads ACCELERATED / EVOLUTION and every SVG aria-label mirrors its markdown alt verbatim.
  • Retitles the previous post to "Fighting AI Training Bias" (frontmatter, banner alt, SVG aria-label; the drawn headline already matched) and points its sequel line at the new post. Slug and canonical URL unchanged.
  • Adds a blog voice rule to docs/blog/CLAUDE.md: every post stands alone for a stranger — repo artifacts are link receipts behind descriptive words, never narrative glue.

Review findings addressed

  • Custom-CSS table row corrected against eval/skill-loop/runs/*/metrics.json: "56 in the first, then 10 to 21, settling near the ~10-line pattern the skills sanction."
  • Em-dash ExampleMeta excerpt: the code fence is gone entirely — the captured-runs section is prose now, which also fit the story-first rewrite better than doctoring a verbatim quote would have.

Where this deviates from issue #380

Verification

  • pnpm format:check, pnpm lint, and the docs build (onBrokenLinks: 'throw' link validation) all pass.
  • build/.devto-publish/2026-08-06-ai-loop-keeps-skills-honest.md regenerated with production-URL rewrites and no code fences; dev.to/Medium publishing stays manual post-merge.
  • Built output: og:title and sidebar show the new title first, the retitled bias post second; the post file carries today's date per the filename convention.

Closes #380

Summary by CodeRabbit

  • New Features

    • Added a new blog post covering agent skills, evaluation, iterative improvements, demonstrations, safeguards, and experiment limitations.
  • Documentation

    • Updated a related article’s title, image description, and sequel link.
    • Added guidelines requiring blog posts to provide standalone context and use descriptive links.

Copilot AI balanced review requested due to automatic review settings August 5, 2026 03:54

Copilot AI left a comment

Copy link
Copy Markdown
Contributor

Choose a reason for hiding this comment

The reason will be displayed to describe this comment to others. Learn more.

Copilot was unable to review this pull request because the user who requested the review has reached their quota limit.

@coderabbitai

coderabbitai Bot commented Aug 5, 2026 •

Copy link
Copy Markdown

Review Change Stack

Warning

Review limit reached

@allxsmith, you've reached your PR review limit, so we couldn't start this review.

Next review available in: 19 minutes

You've used all free OSS reviews for now. Wait for the free limit to reset to keep reviewing this public repository.

How can I continue?

After more reviews become available, a review can be triggered using the @coderabbitai review command as a PR comment. Alternatively, push new commits to this PR.

To avoid repeated limits, reduce automatic review volume by pausing incremental auto-reviews earlier, using label-based review opt-in, excluding WIP or generated PR titles, or requesting reviews manually when the PR is ready. If your team needs uninterrupted high-volume reviews, an organization admin can enable usage-based reviews.

How do review limits work?

CodeRabbit enforces per-developer PR review limits for each organization. Most developers receive the normal plan review availability.

For paid Pro and Pro+ PR reviews, CodeRabbit uses adaptive limits for sustained high-volume activity. When a developer's recent PR review activity reaches the 95th percentile or higher among CodeRabbit users, additional reviews become available more gradually as earlier reviews age out of the rolling window.

Please refer docs for additional details.

Review details
⚙️ Run configuration

Configuration used: Path: .coderabbit.yaml

Review profile: CHILL

Plan: Pro Plus

Run ID: 8bee000f-0ef1-447c-b292-54775cfaaed2

📥 Commits

Reviewing files that changed from the base of the PR and between f896edd and deb7da4.

📒 Files selected for processing (1)
  • docs/blog/2026-08-06-ai-loop-keeps-skills-honest.md

Walkthrough

Added a blog post about agent-skill evaluation and maintenance. Updated a related post’s title, alt text, and sequel link. Added conventions for standalone blog posts and descriptive repository links.

Changes

AI skills documentation

Layer / File(s) Summary
Evaluation workflow
docs/blog/2026-08-06-ai-loop-keeps-skills-honest.md
Added the article front matter, background, cold-start evaluation procedure, ten-generation improvement loop, scorecards, metrics, and experimental conditions.
Evidence and review gates
docs/blog/2026-08-06-ai-loop-keeps-skills-honest.md
Documented Storybook skill showcases, replay metadata, adversarial review, human approval, CI checks, findings, and experiment limitations.
Companion article alignment
docs/blog/2026-08-05-fighting-ai-training-bias.md, docs/blog/CLAUDE.md
Shortened the companion article title and alt text, linked its published sequel, and added standalone-post writing conventions.

Estimated code review effort: 1 (Trivial) | ~5 minutes

Possibly related issues

Possibly related PRs

Suggested labels: documentation

Suggested reviewers: copilot

🚥 Pre-merge checks | ✅ 5
✅ Passed checks (5 passed)
Check name Status Explanation
Title check ✅ Passed The title clearly identifies the new skills blog post and the related retitle, which are the main changes in the pull request.
Description check ✅ Passed The description gives a detailed summary, affected docs, issue link, rationale, verification results, and context, although it omits some template checkboxes.
Linked Issues check ✅ Passed The new post, companion retitle, cross-links, safeguards, current evaluation-harness context, and validation satisfy the coding-related objectives in [#380].
Out of Scope Changes check ✅ Passed All reviewed changes support the requested blog post, companion retitle, or standalone blog-writing guidance, with no unrelated code or product changes.
Docstring Coverage ✅ Passed No functions found in the changed files to evaluate docstring coverage. Skipping docstring coverage check.
✨ Finishing Touches
🧪 Generate unit tests (beta)
  • Create PR with unit tests
  • Commit unit tests in branch docs/380-ai-loop-keeps-skills-honest

Thanks for using CodeRabbit! It's free for OSS, and your support helps us grow. If you like it, consider giving us a shout-out.

❤️ Share

Comment @coderabbitai help to get the list of available commands.

@github-actions

github-actions Bot commented Aug 5, 2026

Copy link
Copy Markdown
Contributor

Preview Deployment

Preview URL: https://f9139633.bestax.pages.dev

@coderabbitai coderabbitai Bot left a comment

Copy link
Copy Markdown

Choose a reason for hiding this comment

The reason will be displayed to describe this comment to others. Learn more.

Actionable comments posted: 1

🤖 Prompt for all review comments with AI agents
Verify each finding against current code. Fix only still-valid issues, skip the
rest with a brief reason, keep changes minimal, and validate.

Inline comments:
In `@docs/blog/2026-08-05-ai-loop-keeps-skills-honest.md`:
- Line 79: Update the prompt literal in the example to replace both em dashes
with blog-approved punctuation, preserving the prompt’s meaning and wording.
🪄 Autofix

Fix all unresolved CodeRabbit comments on this PR:

  • Push a commit to this branch (recommended)
  • Create a new PR with the fixes

ℹ️ Review info
⚙️ Run configuration

Configuration used: Path: .coderabbit.yaml

Review profile: CHILL

Plan: Pro Plus

Run ID: b646f7a7-f24b-4071-b4a2-9fa2b831f5e7

📥 Commits

Reviewing files that changed from the base of the PR and between 54b78cd and 1b988b0.

⛔ Files ignored due to path filters (15)
  • docs/static/img/ai-loop-keeps-skills-honest-captured-runs.png is excluded by !**/*.png
  • docs/static/img/ai-loop-keeps-skills-honest-captured-runs.svg is excluded by !**/*.svg
  • docs/static/img/ai-loop-keeps-skills-honest-findings.png is excluded by !**/*.png
  • docs/static/img/ai-loop-keeps-skills-honest-findings.svg is excluded by !**/*.svg
  • docs/static/img/ai-loop-keeps-skills-honest-gates.png is excluded by !**/*.png
  • docs/static/img/ai-loop-keeps-skills-honest-gates.svg is excluded by !**/*.svg
  • docs/static/img/ai-loop-keeps-skills-honest-loop.png is excluded by !**/*.png
  • docs/static/img/ai-loop-keeps-skills-honest-loop.svg is excluded by !**/*.svg
  • docs/static/img/ai-loop-keeps-skills-honest-proof.png is excluded by !**/*.png
  • docs/static/img/ai-loop-keeps-skills-honest-proof.svg is excluded by !**/*.svg
  • docs/static/img/ai-loop-keeps-skills-honest-shipped.png is excluded by !**/*.png
  • docs/static/img/ai-loop-keeps-skills-honest-shipped.svg is excluded by !**/*.svg
  • docs/static/img/ai-loop-keeps-skills-honest.png is excluded by !**/*.png
  • docs/static/img/ai-loop-keeps-skills-honest.svg is excluded by !**/*.svg
  • docs/static/img/fighting-ai-training-bias.svg is excluded by !**/*.svg
📒 Files selected for processing (2)
  • docs/blog/2026-08-05-ai-loop-keeps-skills-honest.md
  • docs/blog/2026-08-05-fighting-ai-training-bias.md

Comment thread docs/blog/2026-08-05-ai-loop-keeps-skills-honest.md Outdated
Comment thread docs/blog/2026-08-05-ai-loop-keeps-skills-honest.md Outdated

@claude claude Bot left a comment

Copy link
Copy Markdown

Choose a reason for hiding this comment

The reason will be displayed to describe this comment to others. Learn more.

Deep review — 1 blocking · 1 advisory

# Severity Area Finding Location
1 🟡 Minor Correctness Custom-CSS "10 to 21" excludes i02 (=56); actual i02–i10 range is 10 to 56 docs/blog/2026-08-05-ai-loop-keeps-skills-honest.md:54
2 🔵 Advisory Robustness Em dashes in the ExampleMeta prompt quote technically fall under the blog's no-em-dash-in-demo-strings rule docs/blog/2026-08-05-ai-loop-keeps-skills-honest.md:79

Overall: Solid, well-sourced docs PR. I verified the headline numbers against eval/skill-loop/report.md (baseline 85 → mean 95.2, median 96, min 89, max 99; -43% cost, -38% turns; 42 raw classNames → 0; 3 grader errors), the rubric's 50-point do-nothing gate against rubric.md, "87 documented components" against the catalog, the "Skills sync (same PR, always)" heading in CONTRIBUTING-COMPONENTS.md, and the skills/CLAUDE.md quotes — all check out. Every issue/PR reference (#367-371, #194-197, #302/#303/#326/#329, #380) resolves to a title matching its description. All 7 SVG+PNG hero pairs exist at 1200×630, each SVG's aria-label mirrors its markdown alt text verbatim, the banner embeds the SVG and section images embed PNGs per the cover contract, and the syndication contract holds (flat .md, /img/-rooted images, no admonitions, the only JSX is inside a tsx fence). The one real defect is a mislabeled range in the results table; the human should focus there.

Residual risk:

  • Build/link validation — I could not run pnpm build/format:check in this sandbox (pnpm/npx blocked), but I manually confirmed every internal link target exists (/docs/skills/intro, /docs/guides/getting-started/ai-development, both /blog/ slugs), so onBrokenLinks: 'throw' should pass; CI still runs the real build.
  • Other tabular numbers — spot-checked the full report row-by-row; every other cell (baseline, mean, median/min/max, cost, turns, classNames) matches the source. Only the custom-CSS range was off.
  • Date-vs-filename convention — the 2026-08-05- prefix matches today; if merge slips past that date the file needs a rename per blog/CLAUDE.md, handled at merge time.

🏄 Clean set, dude — the whole thing rides on receipts and almost all of 'em check out. Just one number wiped out on the CSS wave (56 snuck into a 10-to-21 lineup). Patch that cell and it's all-time, go ahead and paddle it in.

Copilot AI review requested due to automatic review settings August 7, 2026 00:14

Copilot AI left a comment

Copy link
Copy Markdown
Contributor

Choose a reason for hiding this comment

The reason will be displayed to describe this comment to others. Learn more.

Copilot was unable to review this pull request because the user who requested the review has reached their quota limit.

@allxsmith allxsmith changed the title docs: add the AI loop keeps the skills honest post and retitle the bias post docs: add the accelerated evolution skills post and retitle the bias post Aug 7, 2026
@allxsmith

Copy link
Copy Markdown
Owner Author

Rewrote the post top to bottom as a self-contained story ("Accelerated Evolution for Agent Skills") — no series numbering, no internal file/check names as narrative glue, project introduced in a clause, and the bias argument retold in two sentences for readers landing cold from dev.to or Medium.

Both review findings are addressed in the rewrite: the custom-CSS row now matches the committed run metrics (56 in the first revised run, then 10 to 21), and the ExampleMeta code excerpt is gone entirely (the captured-runs section is prose), which moots the em-dash thread without editing a verbatim quote of shipped code.

Also in this push: the post file's date bump to 2026-08-06, a regenerated cover headline (ACCELERATED / EVOLUTION; scene unchanged), a reworded sequel line in the bias post, and a new blog voice rule in docs/blog/CLAUDE.md codifying the stand-alone-for-strangers requirement.

@github-actions

github-actions Bot commented Aug 7, 2026

Copy link
Copy Markdown
Contributor

Preview Deployment

Preview URL: https://605893bd.bestax.pages.dev

@coderabbitai coderabbitai Bot left a comment

Copy link
Copy Markdown

Choose a reason for hiding this comment

The reason will be displayed to describe this comment to others. Learn more.

Actionable comments posted: 2

🤖 Prompt for all review comments with AI agents
Verify each finding against current code. Fix only still-valid issues, skip the
rest with a brief reason, keep changes minimal, and validate.

Inline comments:
In `@docs/blog/2026-08-06-ai-loop-keeps-skills-honest.md`:
- Line 97: Rewrite the closing sentence in the paragraph so it clearly states
that running the evolution loop quickly keeps improving the skills, using the
requested wording and preserving the surrounding content.
- Line 65: Update the sentence in the paragraph beginning “And that's what
‘accelerated’ means here” to replace the ungrammatical training-data time
reference with clear wording that it reaches agents only in a future model.
Preserve the surrounding meaning and proofread the paragraph for clean
mechanics.
🪄 Autofix

Fix all unresolved CodeRabbit comments on this PR:

  • Push a commit to this branch (recommended)
  • Create a new PR with the fixes

ℹ️ Review info
⚙️ Run configuration

Configuration used: Path: .coderabbit.yaml

Review profile: CHILL

Plan: Pro Plus

Run ID: 765ae25d-9961-4da6-8cd9-f81a475257e0

📥 Commits

Reviewing files that changed from the base of the PR and between 1b988b0 and f896edd.

⛔ Files ignored due to path filters (2)
  • docs/static/img/ai-loop-keeps-skills-honest.png is excluded by !**/*.png
  • docs/static/img/ai-loop-keeps-skills-honest.svg is excluded by !**/*.svg
📒 Files selected for processing (3)
  • docs/blog/2026-08-05-fighting-ai-training-bias.md
  • docs/blog/2026-08-06-ai-loop-keeps-skills-honest.md
  • docs/blog/CLAUDE.md
🚧 Files skipped from review as they are similar to previous changes (1)
  • docs/blog/2026-08-05-fighting-ai-training-bias.md

Comment thread docs/blog/2026-08-06-ai-loop-keeps-skills-honest.md Outdated
Comment thread docs/blog/2026-08-06-ai-loop-keeps-skills-honest.md Outdated
Copilot AI review requested due to automatic review settings August 7, 2026 00:39

Copilot AI left a comment

Copy link
Copy Markdown
Contributor

Choose a reason for hiding this comment

The reason will be displayed to describe this comment to others. Learn more.

Copilot was unable to review this pull request because the user who requested the review has reached their quota limit.

@github-actions

github-actions Bot commented Aug 7, 2026

Copy link
Copy Markdown
Contributor

Preview Deployment

Preview URL: https://fc76e517.bestax.pages.dev

@allxsmith
allxsmith requested a balanced review from Copilot August 7, 2026 00:45

Copilot AI left a comment

Copy link
Copy Markdown
Contributor

Choose a reason for hiding this comment

The reason will be displayed to describe this comment to others. Learn more.

Copilot was unable to review this pull request because the user who requested the review has reached their quota limit.

@allxsmith

Copy link
Copy Markdown
Owner Author

@coderabbitai review

@coderabbitai

coderabbitai Bot commented Aug 7, 2026 •

Copy link
Copy Markdown
⚠️ Action not completed

No files to review.

Note: CodeRabbit is an incremental review system and does not re-review already reviewed commits. This command is applicable only when automatic reviews are paused.

@claude claude Bot left a comment

Copy link
Copy Markdown

Choose a reason for hiding this comment

The reason will be displayed to describe this comment to others. Learn more.

Deep review — 0 blocking · 1 advisory

# Severity Area Finding Location
1 🔵 Advisory Correctness "One npx skills add command drops them into any existing project" reads as installing all seven at once, but the shipped path is one npx skills add … --skill <name> invocation per skill (skills/README.md:27-33); the single-command install is npm create bestax, which the sentence already covers separately. docs/blog/2026-08-06-ai-loop-keeps-skills-honest.md:27

Overall: This is a sound, carefully-sourced docs PR (one new blog post, a retitle of the prior post, and 7 SVG+PNG image pairs). I chased every quantitative and reference claim to its source and they hold: the score table (85 -> mean 95.2, median 96 / min 89 / max 99; raw CSS 42->0; custom CSS 77->56->range 10-21; $10.55/127 -> mean $6.00/79) matches eval/skill-loop/report.md exactly, the "+/-6 point single-run swing on identical tooling" limit matches report.md:227 and README.md:60, and all twelve issue/PR references (#367-#371, #380, #382, #384, #196, #302/#303/#329) resolve and describe what the prose says. Nothing structural is at risk, just the one prose generalization about the install command above.

Residual risk: the failure class here is inaccurate or broken published content.

  • Broken links - refuted: every internal target (/docs/skills/intro, /docs/guides/getting-started/ai-development, /blog/fighting-ai-training-bias, all /img/*) resolves to a committed source file, and onBrokenLinks: 'throw' gates the build; external links are GitHub/Storybook absolute URLs.
  • Accessibility regressions - refuted: all 7 new SVG aria-labels mirror their markdown alt verbatim, and the retitled bias SVG's aria-label was updated in lockstep with its alt/title.
  • Syndication breakage - refuted: the publish_to_devto: true post is clean plain markdown (no JSX, no ::: admonitions), ships PNG heroes plus an SVG banner per convention, carries the <!-- truncate --> fold, and prettier-check passes.

🏄 Total cruiser of a PR, dude - every number in the set actually catches the wave it claims, the alt-text and aria are riding in perfect sync, and the only ripple is one sentence that makes skills add sound like a one-and-done. Clean it or don't, it is still good to paddle out.

Copilot AI review requested due to automatic review settings August 7, 2026 00:55

Copilot AI left a comment

Copy link
Copy Markdown
Contributor

Choose a reason for hiding this comment

The reason will be displayed to describe this comment to others. Learn more.

Copilot was unable to review this pull request because the user who requested the review has reached their quota limit.

@allxsmith
allxsmith merged commit 6b57479 into main Aug 7, 2026
51 checks passed
@allxsmith
allxsmith deleted the docs/380-ai-loop-keeps-skills-honest branch August 7, 2026 00:56
@github-actions

github-actions Bot commented Aug 7, 2026

Copy link
Copy Markdown
Contributor

Preview Deployment

Preview URL: https://3b77a0eb.bestax.pages.dev

@bestax-release-bot

Copy link
Copy Markdown

🎉 This PR is included in version 5.8.1 🎉

The release is available on:

Your semantic-release bot 📦🚀

@bestax-release-bot

Copy link
Copy Markdown

🎉 This PR is included in version 4.0.2 🎉

The release is available on:

Your semantic-release bot 📦🚀

@bestax-release-bot

Copy link
Copy Markdown

🎉 This PR is included in version 1.0.0 🎉

The release is available on:

Your semantic-release bot 📦🚀

@bestax-release-bot

Copy link
Copy Markdown

🎉 This PR is included in version 2.0.1 🎉

The release is available on:

Your semantic-release bot 📦🚀

Sign up for free to join this conversation on GitHub. Already have an account? Sign in to comment

Projects

None yet

Development

Successfully merging this pull request may close these issues.

[Docs] Blog: How our AI loop keeps the skills honest

2 participants