Repository navigation
Add a 32-layer landscape-sweep lane to manifest-20260923 and source reviews for its newcomers - #153
Conversation
…osals, 11 survivors with source reviews A per-layer sweep of all 32 layers (one researcher per layer, a facts/identity and a fit/standing refuter per layer, every worker at effort max) proposed 98 repositories the catalog did not know; 11 withstood both refuters. The manifest is rebuilt with the unchanged generator from the original private inputs plus this lane (the original inputs alone reproduce the committed manifest byte for byte). Survivors carry a neutral source_review file at a pinned upstream commit, registered and listed first in their evidence, for the blind re-record lanes. Co-Authored-By: Claude Opus 5.5 (1M context) <noreply@anthropic.com>
Codex Review SummaryThis comment shows the latest Codex review activity on this pull request.
ℹ️ About Codex in GitHubYour team has set up Codex to review pull requests in this repo. Reviews are triggered when you
Codex reacts with 👀 while any review is running, comments if it has suggestions, and reacts with 👍 once all reviews finish with no findings. |
There was a problem hiding this comment.
💡 Codex Review
Here are some automated review suggestions for this pull request.
Reviewed commit: bf08e09581
ℹ️ About Codex in GitHub
Your team has set up Codex to review pull requests in this repo. Reviews are triggered when you
- Open a pull request for review
- Mark a draft as ready
- Comment "@codex review".
If Codex has suggestions, it will comment; otherwise it will react with 👍.
Codex can also answer questions or update the PR. Try commenting "@codex address that feedback".
…ceipts, existing newcomers covered - F2: repositories already recorded in catalogs/foundation/automation.json (active osv-scanner and Scorecard lanes; dated considered_not_activated decisions for Renovate and gh-aw; harden-runner) are known, not candidates; the two active ones become open gaps on their layers instead. - F1/F3/F4: receipts carry only the upstream's own words (repository description and verbatim README excerpts at the pinned commit) and only the pinned tree/README URLs; popularity, standing, recency, status and comparison sentences are filtered; winner names match whole words; nav-only excerpts and user/sponsor sections are skipped. - agent-lab-17's request: the manifest's 57 existing surviving newcomers (51 repositories, 6 of them Hugging Face models) get the same neutral source-review files, prepended to their evidence. - gitleaks: a rule-scoped, exact-file, whole-line allowlist for the manifest's "pin" git commit ids, which the sourcegraph-access-token rule's keyword activated once lane text named Sourcegraph; with regression tests that other fields and other files stay detected. Co-Authored-By: Claude Opus 5.5 (1M context) <noreply@anthropic.com>
|
Update at
Local checks.
A re-review is running. |
…is a mirror; backtrader unpushed since 2024-08 Co-Authored-By: Claude Opus 5.5 (1M context) <noreply@anthropic.com>
…er, no duplicate backtrader gap
- L1: a sentence dropped between two kept ones is marked [...] so a joined excerpt never reads as
contiguous upstream text.
- L2: counts and superlatives outside the word-boundary filter ("1B+", "#1", "most complete",
"leaderboard") are dropped.
- L3: the backtrader gap is removed; its home layer already carries the beyond lane's
unmaintained_signal.
Pinned commits are unchanged (reused from the reviewed receipts); 10 receipts change in text only.
Co-Authored-By: Claude Opus 5.5 (1M context) <noreply@anthropic.com>
- Codex P1 (RD-Agent): its comparison scores on a final segment the feedback loop never sees. - Codex P2 (TruffleHog, 2 layers): the verification arm uses a controlled, revocable credential. - Codex P2 (MinerU): stale v1.0-era OmniDocBench figures and CJK framing removed. - Codex P2 (Marker): model-weight license terms recorded separately from the code license. - Codex P1 (stopped run): the first run's 9 discovery returns are retained, label-free, with call counts and whether the completed run re-proposed each repository; its usage is unknown. - The completeness critic's follow-up round (same two-refuter rule) is merged: 107 proposals, 11 survivors (adds PaddleOCR, ollama, betterleaks, claude-code-action), with prior documentary records disclosed; completed-run usage recorded (119 agents, 12.1M subagent tokens). Co-Authored-By: Claude Opus 5.5 (1M context) <noreply@anthropic.com>
…tion records disclosed - M1: PaddleOCR-VL-1.6 is 3rd on OmniDocBench v1.6 (README at f133a71e9e), not 1st; its proposed label is lowered to keep_but_compare because its fit vote called the rank-1 basis void. - M2: claude-code-action's prior records are cited correctly (targeted_candidate with an executed CI smoke arm in the 2026-09-22 SDK sweep; decision HOST-09) in the lane limits and its comparison. - L1: RD-Agent keeps its reproduce-the-published-advantage stay condition. - L2/L3: a stale Release claim is corrected and a private scratch name is removed from evidence. Co-Authored-By: Claude Opus 5.5 (1M context) <noreply@anthropic.com>
… decision (HOST-09 is keep-but-compare) Co-Authored-By: Claude Opus 5.5 (1M context) <noreply@anthropic.com>
…h merged evidence (#170) Records only, no gate status change: runtime-target IBKR cites the passed 1.231.0 paper receipt (#147) while rc5 local acceptance stays not_established (blocker nautilus#4983); dashboard checkpoint cites the 32-layer sweep (#153) and the #162 mover research result. Independently reviewed. Co-Authored-By: Claude Opus 5.5 (1M context) <noreply@anthropic.com>
Why. manifest-20260923's discovery was thin: 5 web searches across 20 foundation layers, 6 across 12 trading layers and 7 "beyond", with freshness retained (
fetched_this_run=0). A current SOTA repository that was never a candidate cannot win the 20260923 re-record, however blind its lanes are. Agreed with agent-lab-17, the tooling owner, with a merge cut-off of 2026-09-24T02:00Z.Method. Workflow
wf_38aa6d5d-d6c, every worker at effort max:keep_but_compareortargeted_candidate, never promoted (the generator's PROPOSABLE_LABELS).Result. 98 proposals: 11 survivors and 87 refuted, all published with their votes. The survivors:
document-retrieval: datalab-to/marker, opendatalab/MinerUweb-research: exa-labs/exa-mcp-serverci-supply-chain: ossf/scorecard, renovatebot/renovateobservation-inference: grafana/temposecrets-credentials: trufflesecurity/trufflehoggit-github-automation: github/gh-awresearch-factors-ml: lightgbm-org/LightGBM, microsoft/RD-Agentsecurity-supply-chain: google/osv-scannerBuild.
build_manifest.pyis unchanged. I reproduced the committed manifest byte for byte from the private work dirsota-convergence-20260923(lanes-with-critic.json, the committed citation review, both original--checkout-rootvalues). I then appended this lane (new files beside the originals; the originals are not modified) and rebuilt.lane_calls/lane_limitsfor this lane, andcounts.candidates_total/candidates_by_disposition(73 → 171).Evidence for the blind lanes. Each survivor has a neutral
evidence/artifacts/landscape-sweep-20260923/<owner>-<repo>.json(evidence_class: source_review) containing:Popularity, recency, catalog-status and comparison sentences are filtered out, and the layer's current winners are never named. Each file is registered in
manifests/evidence.jsonfiles[] and listed first in its candidate'sevidence[], so agent-lab-17's--manifest-newcomerspackets attach it.Checks (local, raw).
validate.py,evidence_manifest --check,landscape.pyandbuild_verdicts --checkpass.new_host_grand_list --checkandcomponent_matrix --checkpass.verdict_review_gate --base origin/mainpassed.Limits.
upstream_nowis as observed during the run.🤖 Generated with Claude Code