Skip to content

Add searchable ecosystem manifest and token-efficiency guide - #3

Merged
seathatflowsinourveins merged 1 commit into
mainfrom
codex/ecosystem-html-manifest
Sep 20, 2026
Merged

seathatflowsinourveins merged 1 commit into
mainfrom
codex/ecosystem-html-manifest

Conversation

@seathatflowsinourveins

Copy link
Copy Markdown
Owner

The ecosystem catalog is spread across source cards, component pins and receipts. This adds a self-contained HTML manifest that makes those records searchable and connects native agents, skills, workers, memory, retrieval, evaluation and research in one view.

The explorer preserves the 504-repository / 342-star public index and adds Tavily as a separate dated integration. Token efficiency is a core view: native caching and compaction, scoped retrieval, progressive skill loading, bounded outputs, compact worker handoffs and complete usage accounting. Source review, historical execution, installed files and current-host acceptance remain distinct.

The standard-library builder embeds all data, styles and scripts, includes input hashes, escapes source data and rejects unsafe paths/links. The page makes no background network requests. CI checks that the committed HTML matches its inputs.

Validation:

  • Ten new behavior tests cover source joins/confinement, escaping, evidence boundaries, current integrations, changed inputs and reproducible output.
  • Full repository suite: 283 passed, 38 optional-dependency skips. Publication/catalog validators passed; generated artifact hashes are checked.
  • Isolated browser checks passed for search, star/layer/evidence filters, record dialogs, empty states, navigation and mobile layout. No JavaScript errors, overflow or external background requests were observed.
  • Tavily authentication/live search and fresh Codex/Claude skill metadata discovery passed without inference for the discovery probes. A separate legacy Claude standalone task passed 57 tests; its prior turn-cap failure and usage limits remain visible.
  • Independent source/privacy review passed. Public data contains no credentials, account identifiers or personal paths.

@chatgpt-codex-connector

chatgpt-codex-connector Bot commented Sep 20, 2026 •

Copy link
Copy Markdown

Codex Review Summary

This comment shows the latest Codex review activity on this pull request.

Review Status Commit Review trigger
📝 Code Review ✅ Completed 2026-09-20T03:23:01.431462Z 07a2bb5 PR opened
ℹ️ About Codex in GitHub

Your team has set up Codex to review pull requests in this repo. Reviews are triggered when you

  • Open a pull request for review
  • Mark a draft as ready
  • Comment "@codex review" or "@codex security review".

Codex reacts with 👀 while any review is running, comments if it has suggestions, and reacts with 👍 once all reviews finish with no findings.

@seathatflowsinourveins
seathatflowsinourveins merged commit 1528933 into main Sep 20, 2026
4 checks passed

@chatgpt-codex-connector chatgpt-codex-connector Bot left a comment

Copy link
Copy Markdown

Choose a reason for hiding this comment

The reason will be displayed to describe this comment to others. Learn more.

💡 Codex Review

Here are some automated review suggestions for this pull request.

Reviewed commit: 07a2bb5226

ℹ️ About Codex in GitHub

Your team has set up Codex to review pull requests in this repo. Reviews are triggered when you

  • Open a pull request for review
  • Mark a draft as ready
  • Comment "@codex review".

If Codex has suggestions, it will comment; otherwise it will react with 👍.

Codex can also answer questions or update the PR. Try commenting "@codex address that feedback".

Comment on lines +6 to +8
"components": [
{
"repository": "anthropics/claude-code",

Copy link
Copy Markdown

Choose a reason for hiding this comment

The reason will be displayed to describe this comment to others. Learn more.

P1 Badge Register the new source-review decision array

This new /components collection contains five repository recommendations and acceptance gates, but it is not registered as a supplement in catalogs/us-equities/decision-index.json. Consequently, the validated repository union and catalog lookup omit these current reviews, while the explorer exposes them only through untyped, duplicated annotations. Register this array with scripts/catalog_decisions.py --write --supplement docs/ecosystem/source-review.json#/components and rebuild the dependent artifacts.

AGENTS.md reference: AGENTS.md:L25-L29

Useful? React with 👍 / 👎.

Comment on lines +176 to +178
assigned = [layer["id"] for layer in layers if layer["id"] != "beyond" and (
key in [name.casefold() for name in layer["repositories"]]
or any(word.casefold() in haystack for word in layer["keywords"]))]

Copy link
Copy Markdown

Choose a reason for hiding this comment

The reason will be displayed to describe this comment to others. Learn more.

P2 Badge Match layer keywords as terms instead of substrings

Layer assignment currently tests arbitrary substrings, so short keywords generate incorrect navigation tags. In the committed data, for example, prometheus/prometheus is classified under retrieval because the retrieval keyword rag occurs inside storage, even though its record describes metrics and alerting. Tokenize or boundary-match keywords (while preserving intentional stems such as orchestrat) so layer counts and filters do not return unrelated repositories.

Useful? React with 👍 / 👎.


def render(root):
data = build_data(root)
template = safe_file(root, TEMPLATE).read_text(encoding="utf-8")

Copy link
Copy Markdown

Choose a reason for hiding this comment

The reason will be displayed to describe this comment to others. Learn more.

P2 Badge Include the template in embedded provenance

The template is read only after build_data() has finalized the inputs array, so docs/ecosystem/template.html never appears in the explorer's embedded provenance. When its CSS, application script, or CSP changes and the page is rebuilt, the page still claims a complete set of public input hashes while omitting the source that controls its rendering and security behavior. Hash and track the raw template before serializing the provenance data.

Useful? React with 👍 / 👎.

Comment on lines +146 to +148
depth = " ".join(fields[name] for name in (
"evidence_level", "review_level", "review_depth", "evidence_depth"))
reviewed = reviewed or "source_review" in depth or "primary_source" in depth

Copy link
Copy Markdown

Choose a reason for hiding this comment

The reason will be displayed to describe this comment to others. Learn more.

P1 Badge Recognize the repository's declared source-review levels

The review classifier recognizes only strings containing source_review or primary_source, but registered records also use review levels such as source_sections_reviewed and selected_primary_files. As a result, the committed explorer marks reviewed repositories including duckdb/duckdb, openfigi/api-examples, and ariga/atlas as not source-reviewed; records without execution then appear under DISCOVERY / METADATA, and the Source reviewed filter excludes them. Classify the repository's declared review-level vocabulary explicitly instead of relying on these two substrings.

AGENTS.md reference: AGENTS.md:L25-L29

Useful? React with 👍 / 👎.

Comment on lines +67 to +72
def stamp(value):
for key in ("retrieved_at", "checked_at", "recorded_at_utc", "recorded_at",
"observed_at_utc", "observed_at", "executed_at", "date"):
if isinstance(value.get(key), str):
return value[key]
return "Date not recorded in this source"

Copy link
Copy Markdown

Choose a reason for hiding this comment

The reason will be displayed to describe this comment to others. Learn more.

P2 Badge Read the timestamp fields used by canonical receipts

The timestamp extractor omits field names already used by canonical receipts, including observed_utc, started_at_utc, completed_at_utc, observed_date, and observed_date_utc. The generated explorer therefore displays Date not recorded in this source for ten dated receipt IDs—such as the Claude convergence, Alpaca historical, and identity-readiness receipts—even though their exact timestamps are present. Add the existing receipt timestamp variants or normalize receipt dates through a shared schema.

Useful? React with 👍 / 👎.

Sign up for free to join this conversation on GitHub. Already have an account? Sign in to comment

Labels

None yet

Projects

None yet

Development

Successfully merging this pull request may close these issues.

1 participant