diff --git a/README.md b/README.md index 09df2c9fd..389fe3819 100644 --- a/README.md +++ b/README.md @@ -4,10 +4,17 @@ A current, evidence-backed reference for native **Codex + Claude Code**, with sc **Snapshot: September 20, 2026.** This repository records a tested selection and its limits. It is not a claim that every available framework is installed or that an independent universal SOTA benchmark has been won. +The new **[evidence-led convergence practice](blueprints/convergence-practice/README.md)** +connects public own repositories, stars and curated-list discovery to pinned +source review, frozen experiments and explicit adoption decisions. It adds an +executed external retrieval comparison and a reusable acceptance protocol for +native agent work, recovery and local inference. Research findings remain +separate from a new host's runtime acceptance. + The **[US-equities grand catalog](catalogs/us-equities/README.md)** now covers **147 unique repositories in 152 layer decision cards**, **20 model entries**, and an auditable **342-star coverage ledger**. Its combined index includes -**502 repository identities** across all 342 public stars and 160 beyond them, +**504 repository identities** across all 342 public stars and 162 beyond them, with typed, validated pointers to decisions and evidence. The latest [architecture research wave](catalogs/us-equities/architecture/README.md) adds source reviews, awesome-list coverage and official Alpaca constraints. diff --git a/blueprints/convergence-practice/README.md b/blueprints/convergence-practice/README.md new file mode 100644 index 000000000..424f6d93d --- /dev/null +++ b/blueprints/convergence-practice/README.md @@ -0,0 +1,153 @@ +# Evidence-led convergence practice + +This practice turns repository discovery into a tested change to a native agent +workflow. Its target is reliable issue-to-result work across a coordinator and +explicit worker hosts, with source-grounded research, bounded context, native +authentication and recoverable execution. + +The September 20, 2026 wave builds on the existing +[native stack](../../README.md), [source-review union](../../catalogs/us-equities/decision-index.json), +[research runtime](../us-equities/research-runtime/README.md) and +[measured retrieval baseline](../us-equities/retrieval-evaluation/README.md). +It does not replace those receipts or transfer their acceptance to a new host. +SOTA is the research direction; this publication makes no universal best-stack +or end-to-end autonomous-system claim. + +## This wave's evidence + +The [source review](../../catalogs/convergence-practice/source-review.md) refreshed +one public owned repository and all **342 public stars**. All stars were already +represented in the prior public index. Three pinned awesome lists supplied +selected discovery links; seven candidates received pinned README, license and +supporting-source review. The ledger retains 26 selected source-file hashes. +Harbor is a new candidate; llama.cpp was already a maintained trial and is newly +represented in the public index. Neither becomes an installed default here. + +The [machine-readable source ledger](../../catalogs/convergence-practice/source-review.json) +distinguishes the source-review wave from subsequent execution. Its public +inventory is allowlisted metadata; none of the private ecosystem's source, +histories, accounts or machine configuration is included. + +The [external retrieval experiment](arb-trace2code/README.md) runs two unmodified +upstream baselines against a complete published failure-trace subset. Its inputs, +exact implementation pin, observed results and replay instructions are retained. +It evaluates file retrieval, not native agent repair, answer correctness or +provider savings. + +The [local acceptance fixture](local-fixture/fixture.json) freezes 24 questions +and 15 public documents at an immutable repository revision: 20 answerable +questions plus four corpus-specific no-answer cases. Its +[freeze receipt](local-fixture/freeze-receipt.json) predates ranking. The author +read the sources, so this is a transparent integration fixture, not blinded +external relevance judgment. Preserve these exact bytes for any comparison. +The [executed local replay](local-fixture/evaluation.md) retains all 48 rankings, +the denominator rules and a fixture-specific recipe using unchanged upstream APIs. + +## What the comparisons changed + +The external failure-trace subset favors ARB's path/symbol-aware lexical +implementation on mean Recall@20: **69.64% versus 49.34%** for its BM25 baseline. +The separate local documentation fixture favors BM25 on exact first-source +retrieval: **95% versus 80%** across the 20 answerable questions. Both methods +reach 100% Recall@3 on that small local fixture. These are different corpora, +labels and cutoffs; their scores must not be pooled into one leaderboard. + +Both local rankers also return documents for all four no-answer questions. +That observation is not an answer-generation failure rate: no answer model or +abstention policy was evaluated. It demonstrates why a ranking alone cannot +certify that the requested information exists in the corpus. + +The resulting decision is to retain the native production choices and adopt +the replayable evaluation practice. There is no evidence here for replacing +QMD, SocratiCode, memory or a native agent harness. A proposed replacement must +first beat the relevant current workflow on its intended task and pass its +integration and recovery checks. + +## The reusable loop + +1. **Name the failure and the useful outcome.** Start from a real task, a measured + miss or an explicit missing capability. Define success before selecting a tool. +2. **Discover broadly, review narrowly.** Refresh public own repositories and + stars, follow relevant curated-list links, then inspect primary upstream + documentation and source. Record how each candidate was discovered. An awesome + list is a discovery source, not proof of compatibility or quality. +3. **Challenge the current choice.** Compare the existing native workflow, + a minimal alternative and the proposed addition. State what each overlaps, + its license at the exact pin, hardware/dependency needs and expected benefit. +4. **Freeze the experiment.** Pin source and input hashes, questions, relevance + judgments, budgets and evaluation rules before running. A question authored + from source is a repository-specific fixture, not an independent user study. +5. **Run in isolation.** Use an explicit project and disposable environment. + Retain successful results, failures and skipped work. Keep credentials in + native stores and raw machine details out of the publication. +6. **Decide from the result.** Adopt only the capability and host scope actually + accepted. Retain the current implementation when the candidate has no measured + advantage. An inconclusive result is useful evidence and leaves the gate open. +7. **Publish and replay.** Include commands, input/code hashes, metric definitions, + limitations, rollback and the next unresolved test. Check the exact publication + commit in CI. Revisit after a relevant upstream, model, corpus or host change. + +[protocol.json](protocol.json) makes the lanes and required evidence explicit. +The contract adds no scheduler, always-loaded instructions, proxy, memory store or +runtime service. Model calls, package installation and broker activity remain +separate actions with their own scope. + +## Keep three decisions separate + +| Question | Evidence that can answer it | +| --- | --- | +| Is the project worth investigating? | Public discovery provenance and primary-source review | +| Does this pinned implementation work here? | A useful execution on the named platform, with inputs and observed output | +| Does it improve our work? | A matched task comparison with correctness, resource use and failure behavior | + +Discovery, source review, offline artifact checks, native CLI behavior, +model-mediated task outcomes and recovery acceptance are distinct evidence +classes. They are not interchangeable badges. Record an adoption decision +separately: **retain**, **trial**, **adopt within scope**, **defer**, or **reject**. +State which fact would change that decision. + +## Next acceptance lanes + +| Lane | Small useful acceptance | What remains unproved by a version check | +| --- | --- | --- | +| Retrieval | Frozen source questions, exact identifiers, negative cases, and a larger external benchmark | Answer correctness, safe abstention and successful code edits | +| Native agent engineering | One real issue, isolated change, relevant tests and independently reviewed patch | General autonomous repair or production safety | +| Memory | Retrieve an explicitly approved decision in its project and refuse a different scope | Automatic consolidation quality or safe transcript capture | +| Access and recovery | Reconnect after interruption, preserve job identity, cancel descendants and recover selected state | Cold boot, network changes and restoration on an independent host | +| Local inference | One defined workload compared with the current baseline; record quality, memory and latency | Superiority from model size, release date or tokens per second alone | +| Research application | Reproduce a source-to-deterministic-result packet with time/provenance checks | Historical data entitlement, strategy validity or order authority | + +The existing financial research path keeps model reasoning separate from numeric, +risk and order state. Its [research protocol](../us-equities/acceptance-wave/research-protocol.md) +and [open gates](../../catalogs/us-equities/convergence-review.json) remain +authoritative for that application; this practice does not advance those gates. + +## Measure the whole task deliberately + +Report retrieval coverage separately from answer or patch correctness. Specify +exact-match versus partial-path scoring, top-k denominators, ties, empty results +and errors. Never silently score a failed search as a correct abstention. + +Record warm/cold conditions, dependency versions and whether timing includes +startup. Selected text bytes, tokenized context, cache hits, provider usage and +billing are different quantities. Claim net token or cost savings only from a +matched task comparison with complete usage categories and comparable quality. + +For new candidate models or retrieval methods, keep a frozen evaluation set +separate from development examples. Once results have influenced changes, that +set is a regression fixture; use fresh questions for the next generalization +claim. Avoid turning repeated review by similar agents into an independent +benchmark or a statistical confidence claim. + +## Public boundary and continuation + +Publish only public source identities, content hashes, sanitized measurements, +reviewed code and reproducible instructions. Private repositories can consume +this practice locally without exporting their source, prompts, accounts, memory +or operational topology. A public star list is a dated discovery snapshot, not +blanket approval to install its contents. + +Continue by selecting one open lane and the smallest experiment that could change +its decision. Keep the accepted implementation and rollback available while a +candidate is tested. A new release triggers inspection; it does not automatically +change a working pin. diff --git a/blueprints/convergence-practice/arb-trace2code/README.md b/blueprints/convergence-practice/arb-trace2code/README.md new file mode 100644 index 000000000..53f4fa292 --- /dev/null +++ b/blueprints/convergence-practice/arb-trace2code/README.md @@ -0,0 +1,132 @@ +# ARB trace2code: exact-release replay + +This is an actual offline replay of the complete **101-case `v2_trace2code`** +release using Agent Retrieval Bench's unmodified lexical and BM25 evaluators. +It is one positive-only failure-trace task, not an evaluation of installed +QMD, rg, SocratiCode, an embedding model, or an agent's repair success. + +| Native ARB ranker | Samples | Skipped | Recall@5 | Recall@10 | Recall@20 | MRR | +|---|---:|---:|---:|---:|---:|---:| +| Lexical | 101 | 0 | 0.343234 | 0.481848 | 0.696370 | 0.207453 | +| BM25 | 101 | 0 | 0.222772 | 0.321782 | 0.493399 | 0.163848 | + +Recall is the arithmetic mean of each sample's fraction of gold file paths +retrieved by the cutoff. It is **not** the fraction of solved cases or hit@k. +MRR uses the first exact gold-file rank across the **full file ranking**, not +only its first twenty files. There are 80 cases with one gold file, 16 with two, +and five with three, for 127 gold-file references. File ranks deduplicate native +chunk rankings by first occurrence of each case-sensitive repository path. + +The seven repository counts are caddy 7, clap 1, etcd 4, gin 56, click 26, +pytest 1 and tokio 6. Gin and Click together contribute 82/101 cases. No +repository balancing, confidence interval, significance test, parameter tuning, +or claim of general superiority is made. Per-repository raw means and paired +differences are retained in [metrics.json](metrics.json). + +## Pins, algorithms and licensing + +- Evaluator: [`v0.2.1`, commit `b487f3866cc13dd971819cb902517a6a50282404`](https://github.com/eyuansu62/agent-retrieval-bench/tree/b487f3866cc13dd971819cb902517a6a50282404). +- Dataset: [revision `5901e1ee3aff048290db72edf9c63bc498b79ea3`](https://huggingface.co/datasets/eyuansu71/agent_retrieval_bench/tree/5901e1ee3aff048290db72edf9c63bc498b79ea3), release `v2_trace2code`. +- Archive SHA-256: `19b252e8cfff42107fedc74005dbb6972f2970af33651ce0c1571546819e41c4`. +- The archive contains 98 frozen base-commit corpora, 24,883 snapshot file + records and 356,074 snapshot chunk records. Repeated files across snapshots + are counted repeatedly. Its compressed size is 39,295,446 bytes and its + extracted file bytes total 300,175,960. +- The evaluator has no mandatory third-party runtime dependency. This run used + a separate stdlib-only Python 3.14.7 environment on macOS arm64. + +ARB lexical scoring combines token document frequency and query frequency with +full-path (+25), basename (+8) and symbol (+5) substring bonuses, then divides +by the square root of the number of unique chunk tokens. BM25 retains the +upstream `k1=1.5`, `b=0.75` and query-frequency weighting. Both tokenize camel-case +split, lowercase ASCII alphanumeric text containing path, symbol, kind and +content. Ties use path then chunk identifier. Neither implementation is an +rg or QMD baseline. [Pinned scoring source](https://github.com/eyuansu62/agent-retrieval-bench/blob/b487f3866cc13dd971819cb902517a6a50282404/src/agent_retrieval_bench/baseline.py). + +The evaluator, benchmark metadata and reports are MIT licensed. Corpus content +retains the licenses of its original repositories; no corpus or query text is +redistributed here. See [upstream data licensing](https://github.com/eyuansu62/agent-retrieval-bench/blob/b487f3866cc13dd971819cb902517a6a50282404/DATA_LICENSE.md). + +The first run used the catalog's retained source snapshot `07014c986f3deadb1548c62b32c0ffbe6a81465d`, +which declares version 0.2.1 but is not the release-tag commit. After verifying +the tag discrepancy, both rankers were rerun at the exact release commit. +The complete evaluator source tree is byte-identical at the two commits +(`0fd0466fb3e24c40c0bdf9468872bea9c9a61fa1`); the differences are documentation +and citation changes. Both executions have identical per-sample detail bytes +and metrics. The initial observation is preserved in the receipt's provenance, +not silently relabeled as an exact-tag run. + +## Retained evidence and limits + +[receipt.json](receipt.json) records source/data pins, parameters, environment, +exit codes, observations, checks and artifact hashes. +[input-files.json](input-files.json) hashes every extracted input file; +[source-files.json](source-files.json) hashes evaluator code and license sources. +[paired-samples.json](paired-samples.json) retains only public sample identifiers, +repository/base commits, gold paths and each ranker's exact gold ranks, allowing +independent Recall/MRR recomputation without republishing queries or code chunks. + +The two `upstream-*-summary.json` files preserve native output unchanged. +**Their legacy `@8k` fields are not tokenizer-measured BCY or token savings.** +The pinned baseline uses Python `len(text)` characters, and accepts an oversized +first chunk, so it does not even enforce a strict 8,000-character cap in that +case. Those legacy fields are excluded from headline metrics. No abstention, +answer correctness, span quality, model inference, provider usage or billing +claim follows from this positive-only file-retrieval experiment. + +The exact-release runs were sequential, single warm observations after the +earlier source-identical run. Timing is informational; it is not a controlled +latency benchmark. Both exact-release commands exit 0; schema validation finds +101 valid samples and zero invalid samples. Fifteen upstream corpus/baseline +tests pass. No runtime/default adoption is implied. + +## Replay the upstream evaluator + +Use a new empty directory and an existing Python 3.14.7 executable as +`ARB_PYTHON`. This recipe creates only that directory's environment/data; it +does not install a global tool, register a service or contact a model provider. + +```sh +git clone https://github.com/eyuansu62/agent-retrieval-bench.git upstream +git -C upstream checkout --detach b487f3866cc13dd971819cb902517a6a50282404 +"$ARB_PYTHON" -m venv --without-pip runtime +mkdir data results +curl --fail --location --max-time 120 \ + --output agent_retrieval_bench_v2_trace2code.tar.zst \ + https://huggingface.co/datasets/eyuansu71/agent_retrieval_bench/resolve/5901e1ee3aff048290db72edf9c63bc498b79ea3/releases/v2_trace2code/agent_retrieval_bench_v2_trace2code.tar.zst +runtime/bin/python - <<'PY' +import hashlib +import tarfile +from pathlib import Path, PurePosixPath + +path = Path("agent_retrieval_bench_v2_trace2code.tar.zst") +with path.open("rb") as handle: + actual = hashlib.file_digest(handle, "sha256").hexdigest() +if actual != "19b252e8cfff42107fedc74005dbb6972f2970af33651ce0c1571546819e41c4": + raise ValueError("Archive checksum mismatch") +with tarfile.open(path, "r:zst") as archive: + members = archive.getmembers() + if len(members) != 120 or sum(item.size for item in members) != 300175960: + raise ValueError("Archive size/member inventory mismatch") + for item in members: + name = PurePosixPath(item.name) + if name.is_absolute() or ".." in name.parts or not (item.isfile() or item.isdir()): + raise ValueError("Unsupported archive member") + archive.extractall("data", members=members, filter="data") +PY +PYTHONPATH=upstream/src runtime/bin/python -m agent_retrieval_bench.cli \ + validate data/benchmark/v2_trace2code/samples.jsonl +for ranker in lexical bm25; do + PYTHONPATH=upstream/src runtime/bin/python -m agent_retrieval_bench.cli \ + eval-baseline data/benchmark/v2_trace2code/samples.jsonl \ + --corpus data/corpus/v2_trace2code --ranker "$ranker" \ + --candidate-filter all_files --no-keep-list --no-progress \ + --out "results/$ranker-summary.json" \ + --details "results/$ranker-details.jsonl" +done +``` + +Run from the directory containing `data`: the released manifest's chunk paths +are relative to that directory. Compare metrics and detail hashes, not elapsed +time fields. Do not substitute `--dry-run`, an answer-only candidate list, or a +sample limit for this complete released-subset replay. diff --git a/blueprints/convergence-practice/arb-trace2code/input-files.json b/blueprints/convergence-practice/arb-trace2code/input-files.json new file mode 100644 index 000000000..c08b042f1 --- /dev/null +++ b/blueprints/convergence-practice/arb-trace2code/input-files.json @@ -0,0 +1,522 @@ +[ + { + "path": "data/benchmark/v2_trace2code/manifest.json", + "bytes": 1006, + "sha256": "2a3e8b636f76da01e8ea0ec04c3324d79762f541137100e3483503d52a0cd4d2" + }, + { + "path": "data/benchmark/v2_trace2code/samples.jsonl", + "bytes": 491534, + "sha256": "9d0ff50155fa4f65c1bbb632abc4f1f11393c50fa7b0a9249e06128b368b6266" + }, + { + "path": "data/benchmark/v2_trace2code/trace2code.jsonl", + "bytes": 491534, + "sha256": "9d0ff50155fa4f65c1bbb632abc4f1f11393c50fa7b0a9249e06128b368b6266" + }, + { + "path": "data/corpus/v2_trace2code/caddyserver__caddy/18ab0f955fc1075d7727c7658dbfb734c673a5c9.chunks.jsonl", + "bytes": 6242734, + "sha256": "5cd07b68ca15a9736c9f1c5dc7df69acb9b4a1bc9a358654b5f01ffe310b0ced" + }, + { + "path": "data/corpus/v2_trace2code/caddyserver__caddy/4f504588669e28373f455b694539475ffe4d2926.chunks.jsonl", + "bytes": 6037978, + "sha256": "19a465ae44188c55123355a92ed4d416153b5fdca69a71a14e424e74ddf9c4e6" + }, + { + "path": "data/corpus/v2_trace2code/caddyserver__caddy/7586e68e273fbe84b588e73c014d15d815413648.chunks.jsonl", + "bytes": 6119230, + "sha256": "bd70b0cf0385e4db743a9d5de3096dd56e3732b19d31531ac26ea47dea157db0" + }, + { + "path": "data/corpus/v2_trace2code/caddyserver__caddy/95941a71e87d6ccd6318bfd95c56929b01edfafd.chunks.jsonl", + "bytes": 5787252, + "sha256": "be31843c055367eef33a1e7a637065bb36fcfc16f0267b3cabb096d73d512007" + }, + { + "path": "data/corpus/v2_trace2code/caddyserver__caddy/aed1af59763d54520a9c72f1fd0222d43904ebfd.chunks.jsonl", + "bytes": 6181294, + "sha256": "7bde7899f3f7fdc24d6733ba833a21f7f88efcfaca27a77555cfb4c68833a5df" + }, + { + "path": "data/corpus/v2_trace2code/caddyserver__caddy/ca0ca67fbdb831c026d334dfd77ecc653f321876.chunks.jsonl", + "bytes": 6077136, + "sha256": "e6f59cd9cbf72217eceb466ed6d21a33c5eebcc826bb8949ce37f1718ba07ee3" + }, + { + "path": "data/corpus/v2_trace2code/caddyserver__caddy/ef496e58ef9e7844cf8f1831030713ee5e9b354b.chunks.jsonl", + "bytes": 6247684, + "sha256": "e0dce10b64e51a10446d86ab6af192fd12a5be93f04d784594045137edff6d22" + }, + { + "path": "data/corpus/v2_trace2code/clap-rs__clap/e82e1edf76bcbddf5fe53428d297520d76a6a300.chunks.jsonl", + "bytes": 6626558, + "sha256": "1072a263bb5a956f56dae389fc512612c60ddf69cccbae365f585d0bb9820254" + }, + { + "path": "data/corpus/v2_trace2code/corpus_manifest.jsonl", + "bytes": 36995, + "sha256": "9fef569ebf06d5a96b788d9b6136dd023a324b3d5e795172faf6f0c696d9498f" + }, + { + "path": "data/corpus/v2_trace2code/etcd-io__etcd/8687c97fa5fc7814bf2deaf4b13e666a1b9a0dda.chunks.jsonl", + "bytes": 14057731, + "sha256": "ffd5a57d72ad4b81a2d9aa846e3d9cb484f4c7ef310d6c2ec210f14e268b2136" + }, + { + "path": "data/corpus/v2_trace2code/etcd-io__etcd/b83b67e483936a94e59a61198934e88bf0ef95b1.chunks.jsonl", + "bytes": 13977948, + "sha256": "a07b90b991c85520c75afcf5dd776b3386659173d19a931f7e1c7259cec027c4" + }, + { + "path": "data/corpus/v2_trace2code/etcd-io__etcd/be3acf4ed017efb7008beb08585d890e5364ce93.chunks.jsonl", + "bytes": 13930270, + "sha256": "7263a6e2437af1726fe6a13157c134a8d92a2797f6eadb9ce0eee5f9cb176101" + }, + { + "path": "data/corpus/v2_trace2code/etcd-io__etcd/c64f8fd68c2e75cc65b88bc0d956547a5f77cab1.chunks.jsonl", + "bytes": 13985180, + "sha256": "47f1a98bb390716650ed210ae9a8d5d159c757788da18274e98160d09eb6988b" + }, + { + "path": "data/corpus/v2_trace2code/gin-gonic__gin/05caa5c00e552a27c3bfb659b99c0d79b81fafe4.chunks.jsonl", + "bytes": 1237802, + "sha256": "460a78fef83bdd6c57ee4fe3ee19441c4cbd9ada613f40cf21deadfcac7c34c4" + }, + { + "path": "data/corpus/v2_trace2code/gin-gonic__gin/0cbb30aa940a643e31874f8ad8e355219d306756.chunks.jsonl", + "bytes": 1154800, + "sha256": "de1eef9717b5b3230d6149106100d2f50878e7d60230e3088fc0179aae6452b7" + }, + { + "path": "data/corpus/v2_trace2code/gin-gonic__gin/16cd8cdd4ef9257b1e86b5119df1536de71c6a87.chunks.jsonl", + "bytes": 1113728, + "sha256": "e70b9a8e6216fefaef3f59ef95ed9dce34a3c6f304fa89bcf89bf5d660f42905" + }, + { + "path": "data/corpus/v2_trace2code/gin-gonic__gin/19b877fa50cbbb9282763099fb177a1e5cc5c850.chunks.jsonl", + "bytes": 1524202, + "sha256": "76274aa52c540a43921f20062de62c4c3ca0c53f91c27855faaa6ed44317004f" + }, + { + "path": "data/corpus/v2_trace2code/gin-gonic__gin/28e57f58b184b2305ace192e02496bb89f6fd8cb.chunks.jsonl", + "bytes": 1345069, + "sha256": "14ab32279d2002077f32d51ffb24fa996eb6cf291e88dd15c14f7ab933a19edd" + }, + { + "path": "data/corpus/v2_trace2code/gin-gonic__gin/2a794cd0b0faa7d829291375b27a3467ea972b0d.chunks.jsonl", + "bytes": 1522499, + "sha256": "716fa902bc4bd9c41c7e48ee68b49b135d52b2c61180047e276e93c2e4e0d02f" + }, + { + "path": "data/corpus/v2_trace2code/gin-gonic__gin/2e22e5085960205fbb11c25776f6ea76b8053253.chunks.jsonl", + "bytes": 1480819, + "sha256": "5e2a81b205362c3cb37582894a167b62cbb8aa92fe147928b0b33913a3ec0a46" + }, + { + "path": "data/corpus/v2_trace2code/gin-gonic__gin/2e2bd1f408fdaa5e522bc56972e3f109fd7502fd.chunks.jsonl", + "bytes": 1390474, + "sha256": "f88165e7e4e55889e84933548d026e82aa0dc534ee9003cc25a16f32b8bd6e09" + }, + { + "path": "data/corpus/v2_trace2code/gin-gonic__gin/334160bab772f6f93767b870f9d07c176cd4aa2b.chunks.jsonl", + "bytes": 1333293, + "sha256": "9b92d0776d39c717acf805945d7bf6ec1a97eace46dcfe5f3a08bf654e5ee956" + }, + { + "path": "data/corpus/v2_trace2code/gin-gonic__gin/38534e2bf98a06e1f62d6b24384e90b5f78699bf.chunks.jsonl", + "bytes": 1587028, + "sha256": "d5c19ea9239cfdac8389c301d74c2533d92629164af223d8d819835d6a2a4e2d" + }, + { + "path": "data/corpus/v2_trace2code/gin-gonic__gin/3ab698dc5110af1977d57226e4995c57dd34c233.chunks.jsonl", + "bytes": 1529390, + "sha256": "ed2c34eb6af8a079506eccdab7ce827416485117d8e12967b2dd929aede08cf8" + }, + { + "path": "data/corpus/v2_trace2code/gin-gonic__gin/3cb30679b5e3021db16c776ed7e70d380586e9ce.chunks.jsonl", + "bytes": 1343928, + "sha256": "f4f6c1d422d06c09ab2e6ef66bea8eaa643951dbf88d85adbcfea431a895236e" + }, + { + "path": "data/corpus/v2_trace2code/gin-gonic__gin/3f5b0afa2ac85ea79638ca08f4140ce64b8246e5.chunks.jsonl", + "bytes": 1327161, + "sha256": "36f28c5361445fa1598df445672430ba792bacd170a0c8fb5ee26f7ccbc61389" + }, + { + "path": "data/corpus/v2_trace2code/gin-gonic__gin/44d0dd70924dd154e3b98bc340accc53484efa9c.chunks.jsonl", + "bytes": 1271943, + "sha256": "abeeb23947791ffb9ec90a8ba063cb53704242f65777b2e795d49085d45fe857" + }, + { + "path": "data/corpus/v2_trace2code/gin-gonic__gin/472d086af2acd924cb4b9d7be0525f7d790f69bc.chunks.jsonl", + "bytes": 1587050, + "sha256": "dfb432cc410187301f466bfee54d75fa46662a9a12c77e04b74058e35b19bdd7" + }, + { + "path": "data/corpus/v2_trace2code/gin-gonic__gin/4f339e6a35b163d31b30916b37f4176d385f41bd.chunks.jsonl", + "bytes": 1327349, + "sha256": "f8b5d64a1f7be298e35e83ccda1ba7db72cb62d0da88b19bd6e71e6379b89c48" + }, + { + "path": "data/corpus/v2_trace2code/gin-gonic__gin/5e5ff3ace496a31b138b0820136a146bfb5de0ef.chunks.jsonl", + "bytes": 1482198, + "sha256": "05d092369c9d36c80653aa562bf64a01d2bf837bb0378398ee2ba7bc196de91e" + }, + { + "path": "data/corpus/v2_trace2code/gin-gonic__gin/61c2b1c28f0c5a754330545e31f02cd6d6f7944e.chunks.jsonl", + "bytes": 1400343, + "sha256": "9dbd467d8bcc4d471605c9b758d3cd01b4955767eea84b284285104ef98c6dc2" + }, + { + "path": "data/corpus/v2_trace2code/gin-gonic__gin/64ead9e6bd924d431f4dd612349bc5e13300e6fc.chunks.jsonl", + "bytes": 1337139, + "sha256": "4419cf86a2681fcdbc642bf8d746a2097a183ea3b35658d3567b7c79de1c9658" + }, + { + "path": "data/corpus/v2_trace2code/gin-gonic__gin/67c9d4ee110e9adfe33063ef847dba56717c148a.chunks.jsonl", + "bytes": 1390003, + "sha256": "f3d8fdfea3f2f614e0df993b78738704c754a4e8211824975aaa2bb81047172b" + }, + { + "path": "data/corpus/v2_trace2code/gin-gonic__gin/688a429d19d8c804447bb889d3635e2c31a5564d.chunks.jsonl", + "bytes": 1432880, + "sha256": "592a51310a3eb1c28bddcf662b34c52e728ec90f5295f3e6b4fcaeac464c5cfd" + }, + { + "path": "data/corpus/v2_trace2code/gin-gonic__gin/76dd08d512504b80ef76a76c9e6bd1831e121b71.chunks.jsonl", + "bytes": 1433186, + "sha256": "9485670038f79adfc8332ce322c118a99ee093128294f67f76f9cd93964b025e" + }, + { + "path": "data/corpus/v2_trace2code/gin-gonic__gin/79a61b90324586d3c3a5859f8755cae2d1c46f2d.chunks.jsonl", + "bytes": 1259972, + "sha256": "10f31186d63b06a1fb3179f64e5a8dc97d6a731e055857c9cffdfe4617187700" + }, + { + "path": "data/corpus/v2_trace2code/gin-gonic__gin/7a865dcf1dbe6ec52e074b1ddce830d278eb72cf.chunks.jsonl", + "bytes": 1303395, + "sha256": "822f124995d9655ed192330d8bd3445a8e16cf6982a8e001124e0a6c2aad8b9a" + }, + { + "path": "data/corpus/v2_trace2code/gin-gonic__gin/86ff4a64c7efe1a1c875529835eeef9e15de1e86.chunks.jsonl", + "bytes": 1279951, + "sha256": "2755758a921a51861fc51228de2c6d7e0a8aa7cd8ec2c37dca58a397602006da" + }, + { + "path": "data/corpus/v2_trace2code/gin-gonic__gin/90cf4602698dcbce18df3165b2d24e2940670a41.chunks.jsonl", + "bytes": 1375685, + "sha256": "8add2e0a13d3342de7c662108598634ee1377c873a64ddfb51288bda597ddbe0" + }, + { + "path": "data/corpus/v2_trace2code/gin-gonic__gin/915e4c90d28ec4cffc6eb146e208ab5a65eac772.chunks.jsonl", + "bytes": 1528093, + "sha256": "04de6a1d8c0230342b8033b6ee2b9d659c7f08a61a0a183d1c774827fac271a2" + }, + { + "path": "data/corpus/v2_trace2code/gin-gonic__gin/9c081de9cdd1948f521d47d170d18cbc2981c33a.chunks.jsonl", + "bytes": 1337139, + "sha256": "a1ab32d7f9e6c2ef14c9bc142b4ae3aee7c429fe3defaefaafb7ba3d32ed3542" + }, + { + "path": "data/corpus/v2_trace2code/gin-gonic__gin/9f598a31aafb92d675f38f1c8371e4ac76f858bf.chunks.jsonl", + "bytes": 1278519, + "sha256": "9978997b1d71afb13c766d1d19294d08386798864e2c004dc9bd5cc2cbf88e33" + }, + { + "path": "data/corpus/v2_trace2code/gin-gonic__gin/a481ee2897af1e368de5c919fbeb21b89aa26fc7.chunks.jsonl", + "bytes": 1269604, + "sha256": "1e64ef8dac5eacba5c828b013e166245958828d90e7c819357c08e275532b5d5" + }, + { + "path": "data/corpus/v2_trace2code/gin-gonic__gin/a4ac275e079d46d493965491d686bfe72d121e85.chunks.jsonl", + "bytes": 1437960, + "sha256": "be89724490441e11556c9b27056e2b2e9e5878e690b6d6db226aebef76a59d45" + }, + { + "path": "data/corpus/v2_trace2code/gin-gonic__gin/a64286a7760be2031209686ce4d36e99d42dd419.chunks.jsonl", + "bytes": 1275655, + "sha256": "6c0a22aa9715c9cafe304b9cef90e87f3e18a4d9ebe7feef7fe35879aed14a53" + }, + { + "path": "data/corpus/v2_trace2code/gin-gonic__gin/ac95fa6bbcf5ec102bc81fbbb70e8fccba90f0b1.chunks.jsonl", + "bytes": 1558908, + "sha256": "c377c0896e1d06c8dc32d3f613f1fea05200a8ebee552b1466557778daa7591e" + }, + { + "path": "data/corpus/v2_trace2code/gin-gonic__gin/b1c1e7b572f76071fb0e0e7884a0697e0458aa7c.chunks.jsonl", + "bytes": 1307113, + "sha256": "8b4e0482200a55990cea5fb41e4ec1e5bbdb56eb494a91b647ae9e9cfd350211" + }, + { + "path": "data/corpus/v2_trace2code/gin-gonic__gin/b2b489dbf4826c2c630717a77fd5e42774625410.chunks.jsonl", + "bytes": 1530107, + "sha256": "e1e8b435ab6adb82d0d8d864885b52dbfd6cd1509af1683ba563c2aee5192a09" + }, + { + "path": "data/corpus/v2_trace2code/gin-gonic__gin/b4f66e965ba9d60257e0de4c25d4ad4bd6115927.chunks.jsonl", + "bytes": 1303750, + "sha256": "2af328c22b4dc8fcb4bb3ed0637b08d34934527114ad2ed514e1c757a5e47e82" + }, + { + "path": "data/corpus/v2_trace2code/gin-gonic__gin/ba093d19477b896ac89a7fc3246af23d290b8e26.chunks.jsonl", + "bytes": 1570425, + "sha256": "b057ddf72d0e21c7f9c8e1688cb67e74713848194696afd1325672943ac3a62c" + }, + { + "path": "data/corpus/v2_trace2code/gin-gonic__gin/bdc1ad7987b6931801af9d50e2df25667fdfaaf4.chunks.jsonl", + "bytes": 1435513, + "sha256": "51af83bd4dd55d903cad3cc9b3bfbea1fd65b4f3d5edf3aedc1f81d056d9ad00" + }, + { + "path": "data/corpus/v2_trace2code/gin-gonic__gin/c221133ee80c46e3a6c50717ca6f1b41d4ab7711.chunks.jsonl", + "bytes": 1479720, + "sha256": "15e565566bc10b41a35991df1188fa64b4aa2975fdcf2b0a5da4f6caa6d19f24" + }, + { + "path": "data/corpus/v2_trace2code/gin-gonic__gin/c2ba8f19ec19914b73290c53a32de479cd463555.chunks.jsonl", + "bytes": 1269357, + "sha256": "ecfb51bd5997e9f02d127e068cbd61fb588fcec2a01a1598a4e0bab71927d3df" + }, + { + "path": "data/corpus/v2_trace2code/gin-gonic__gin/c358d5656d0feb8b310d4ec379bccde46ccc8cc7.chunks.jsonl", + "bytes": 1514752, + "sha256": "98418a00b7584a5e97f717cc4fa87f52fc0a5f9b6346d11981f5ebebcdccbd75" + }, + { + "path": "data/corpus/v2_trace2code/gin-gonic__gin/c4287b1300363cb3dc2c8408299d3ba6deded485.chunks.jsonl", + "bytes": 1395772, + "sha256": "9912db536b43fb2c3c50182b7bbb770a22fa4b685b7b68cccf72ca2b32c29200" + }, + { + "path": "data/corpus/v2_trace2code/gin-gonic__gin/d4a64265f21993368c90602c18e778bf04ef36db.chunks.jsonl", + "bytes": 1266261, + "sha256": "a5394e529f42f251d80e63f449a60bf4717edeff6faceb8e7842febbe0677a6f" + }, + { + "path": "data/corpus/v2_trace2code/gin-gonic__gin/dab5944a7bca8ae37d947dda02ac591afc1177d3.chunks.jsonl", + "bytes": 1447882, + "sha256": "a64ac8b1052620d2e054a2b3f1932624c13cad9ee35bad2b5c5b582543285178" + }, + { + "path": "data/corpus/v2_trace2code/gin-gonic__gin/db309081bc5c137b2aa15701ef53f7f19788da25.chunks.jsonl", + "bytes": 1572110, + "sha256": "bf9ac6a9d689762efaf613286edcebfda8438aa76770f80b78e4dff5720abdbc" + }, + { + "path": "data/corpus/v2_trace2code/gin-gonic__gin/e0d46ded6cb6974d55a255ab122d1aa6ca0cd60e.chunks.jsonl", + "bytes": 1329345, + "sha256": "cb18bf2e5b17d32af522d6f7d52d9f3f80db3cca5e5dd2f8385d69f9422d1e4f" + }, + { + "path": "data/corpus/v2_trace2code/gin-gonic__gin/e46bd521859fdfc83c508f1d42c92cb7f91e9fcb.chunks.jsonl", + "bytes": 1375784, + "sha256": "2559731114aff1e76ae553017ac85031cc0c13e87383a6b8adcf731085a0e11d" + }, + { + "path": "data/corpus/v2_trace2code/gin-gonic__gin/ecb3f7b5e2f3915bf1db240ed5eee572f8dbea36.chunks.jsonl", + "bytes": 1488304, + "sha256": "189a5ff5ae3f0f9110244382e2217a84345442b45c8a34a711dcb38e4abc47a1" + }, + { + "path": "data/corpus/v2_trace2code/gin-gonic__gin/ecdbbbe9483dd12222f2085f717a2c7cb5ac55fe.chunks.jsonl", + "bytes": 1282718, + "sha256": "de7a2e4a4150ea3fd025c7882274018e00c213510d6eb59b0ac02a1c613f3b1f" + }, + { + "path": "data/corpus/v2_trace2code/gin-gonic__gin/ed150e72544949a96e7eb0b7d1151cf1907068fb.chunks.jsonl", + "bytes": 1473478, + "sha256": "58c01683e3cdd941794e7a8a3c286a02d3c1b01a34f9f426e931f51d190b87f2" + }, + { + "path": "data/corpus/v2_trace2code/gin-gonic__gin/f75144a356e57c95bd21a048f0a40492dcdb33c5.chunks.jsonl", + "bytes": 1283346, + "sha256": "b17d842a55a91c56b4c452b9668019e254150b32e38fc847db733f1bac880bc4" + }, + { + "path": "data/corpus/v2_trace2code/gin-gonic__gin/fad706f1216e6d12bdd51d28d5a40ec27e6c6453.chunks.jsonl", + "bytes": 1520884, + "sha256": "d607086fd24da2150ab8b437cef780affa83bcb5bb936fabc3ce18b2845d485e" + }, + { + "path": "data/corpus/v2_trace2code/gin-gonic__gin/fb2583442c4d9bccb75e6d26f1aa6e7c01950db6.chunks.jsonl", + "bytes": 1580900, + "sha256": "35215d2e2c0c0b8c27a87298b7946bc0282160eb0540b185fb8954e2fbb29267" + }, + { + "path": "data/corpus/v2_trace2code/gin-gonic__gin/fc1c43298de675e5252d0b44f97dc5e204bd4e1e.chunks.jsonl", + "bytes": 1264197, + "sha256": "ee100aaff65cbd4d9e51c199bbfa5c4f5568890e6f38f34a1bddfddc393330f7" + }, + { + "path": "data/corpus/v2_trace2code/pallets__click/011b9f9d190c71310264e6c54bae6259f5e38a9f.chunks.jsonl", + "bytes": 1874486, + "sha256": "7444fb9c64c881f044d0d863de090e46154dd0e56925928c4c12950ac8081407" + }, + { + "path": "data/corpus/v2_trace2code/pallets__click/02046e7a19480f85fff7e4577486518abe47e401.chunks.jsonl", + "bytes": 1744891, + "sha256": "b0a7e514808a002b4f0af3804aeec878f610fe8736e31d4677ae34ff6731ba2d" + }, + { + "path": "data/corpus/v2_trace2code/pallets__click/08ecfd0e7c238a2444886d239de6d0e9d83530ec.chunks.jsonl", + "bytes": 1677530, + "sha256": "89a3bc3b0b728d8975af0dcf8507fc38c013aff2ebd62d41332a2b658980bfb9" + }, + { + "path": "data/corpus/v2_trace2code/pallets__click/0bf333e1b77b4ae0e85d387531b38930fc657f12.chunks.jsonl", + "bytes": 1863084, + "sha256": "8f311103d12b79713ce529f86e3c4542b5b0898b60ade85e5a16c990766dc9aa" + }, + { + "path": "data/corpus/v2_trace2code/pallets__click/1458800409ed12076f18451889b0857db36aa522.chunks.jsonl", + "bytes": 2110778, + "sha256": "daf1f348f1b823cfb34de917781aca312bf4999e066c44a115b58ff2d3ac8f73" + }, + { + "path": "data/corpus/v2_trace2code/pallets__click/1a4d8c1bb1e8f8e214ede7223bd2c05dc2ce006a.chunks.jsonl", + "bytes": 1747391, + "sha256": "62ce6252bf9bc7923a6a757e42c706094f78d7d4dbf1c64267bc04acfbc82dcf" + }, + { + "path": "data/corpus/v2_trace2code/pallets__click/2ed395b0b5ac4d56553ff715335f456f812cdc78.chunks.jsonl", + "bytes": 2007038, + "sha256": "c176a406041bcf64e700a2ac82c1b86325d5f6d517ce08394f9df82200eeef22" + }, + { + "path": "data/corpus/v2_trace2code/pallets__click/380008389b32ca2b7de7c73f670a9eee1562004e.chunks.jsonl", + "bytes": 1770964, + "sha256": "6eb43ab3aa29e1ff4dce1a2133b0ae5171d0c76f4ceeecaa794fd54032046cff" + }, + { + "path": "data/corpus/v2_trace2code/pallets__click/422cb2064eed146dbe03ba3aed5be35daeb70414.chunks.jsonl", + "bytes": 1679913, + "sha256": "65ed55a2fded24be480af12c6bf7aa4f4a27c5366f00adc2835b56702383be06" + }, + { + "path": "data/corpus/v2_trace2code/pallets__click/4271fe283dc9365563aebb369ada8d20eee015a8.chunks.jsonl", + "bytes": 1773570, + "sha256": "710465e2715686e2df7ede2a3d72d14a582468771883afbbae296de73f257415" + }, + { + "path": "data/corpus/v2_trace2code/pallets__click/60cad4ad6583bf885c62d6f4597f3555d1ee1a0f.chunks.jsonl", + "bytes": 2146981, + "sha256": "948ddda40dd6e07b78bd42c4b6c930e2c0cb35f2dc1028d4d4071e53ca6b3a3a" + }, + { + "path": "data/corpus/v2_trace2code/pallets__click/6c7624491911f0b18f869e5d0a989506e7b416f8.chunks.jsonl", + "bytes": 1626246, + "sha256": "82fce15e10bc0cf52add41e081f9ad6ca5b41453fb122a0b81d4f5b0518bb1ce" + }, + { + "path": "data/corpus/v2_trace2code/pallets__click/7d7183604158f064390539d83d4a19a978c6b08a.chunks.jsonl", + "bytes": 2005085, + "sha256": "9644379dff187f17be840345795307a35bc7e1990dd396aa5abc8bcb46de8680" + }, + { + "path": "data/corpus/v2_trace2code/pallets__click/7f7bbe4569ea68e8dabee232eade069ef3310aea.chunks.jsonl", + "bytes": 2034097, + "sha256": "eed6cf02751d6f37a20664abff1016ee00cf0bac47729f4b85ec017c4f59a657" + }, + { + "path": "data/corpus/v2_trace2code/pallets__click/82a83fda1fafd4a43c67755140555bfd5afbd592.chunks.jsonl", + "bytes": 1636815, + "sha256": "d83dbff02bcd5817500552add12de826865e93fc186d016d62be7b6a783acacf" + }, + { + "path": "data/corpus/v2_trace2code/pallets__click/8d3449517a97b6fd457a025bd9c3d2acfd301fa4.chunks.jsonl", + "bytes": 2178287, + "sha256": "c9a6e8ea540107900153cf19f749e70fe71feed12d065c8d12adf01616634afd" + }, + { + "path": "data/corpus/v2_trace2code/pallets__click/a1235aacb1be55dc66ddcfefbf64dec44b6ab54d.chunks.jsonl", + "bytes": 1902205, + "sha256": "8a402992b8f20b90e77dcb3e28c6da5370113183bc10ed684eb80364a79b25b7" + }, + { + "path": "data/corpus/v2_trace2code/pallets__click/a8c05427602d12842c53ce48b03b5d6b5c0c6693.chunks.jsonl", + "bytes": 1756094, + "sha256": "744c9147d61eeebcce695d0ba22e86dafb710419bb15c23892dbf85863729155" + }, + { + "path": "data/corpus/v2_trace2code/pallets__click/b625b9f39082f366dd85ad13024e43c8f5e78b87.chunks.jsonl", + "bytes": 1627110, + "sha256": "76a13cb840e397841ee6e97b1521f3e98df2bc2b44a792dddfe913d4efa927b4" + }, + { + "path": "data/corpus/v2_trace2code/pallets__click/bf3a945b3a220589c67d55dcb73cc1c9a9a175b3.chunks.jsonl", + "bytes": 1614176, + "sha256": "14842fabd79889a92e24a1ac88566ea5cf3cd04263b6846f07e5ba61e9f9a08e" + }, + { + "path": "data/corpus/v2_trace2code/pallets__click/bf9da4838986af3bc09a1b16974d154c61dcbe3f.chunks.jsonl", + "bytes": 1682463, + "sha256": "e5a445343d77dc96a4607d1b2d46d0c844fa6971b0229fddebef76023d39817b" + }, + { + "path": "data/corpus/v2_trace2code/pallets__click/c7e1ba8448cbcb2cdd9c1c7f4a592e863dcc3995.chunks.jsonl", + "bytes": 2158027, + "sha256": "9b903b7864c388466e757abff4b770323cea4cebfc7ee06d45494056f72032ec" + }, + { + "path": "data/corpus/v2_trace2code/pallets__click/c88f333de440c5066ff563f784a6561f4713cf78.chunks.jsonl", + "bytes": 1908417, + "sha256": "2719e1bfc18c77db07752fd689916101b2ef0e4b503210e3796c4ea97bc37ec0" + }, + { + "path": "data/corpus/v2_trace2code/pallets__click/ddede2147c9370ab8275f81db4826b6895d2e7dc.chunks.jsonl", + "bytes": 2185917, + "sha256": "ae8b59d7a14519752207789a2b46c5792dae6860035ed6c895c6e6dcfb216959" + }, + { + "path": "data/corpus/v2_trace2code/pallets__click/fa7f035ab58dc4d9b8818b92fbc6a8f3c3234af6.chunks.jsonl", + "bytes": 2043891, + "sha256": "0c2e0b2b71e314595600d87d48a18b980d9f4816528bacaff889d011ac07e22a" + }, + { + "path": "data/corpus/v2_trace2code/pallets__click/fcd85032cff78aa536a6d2b455fb83bfcc02b228.chunks.jsonl", + "bytes": 1768144, + "sha256": "116dff948cb362813339909f23d43ccb7780db7cfad8029e835d39c64286abc6" + }, + { + "path": "data/corpus/v2_trace2code/pytest-dev__pytest/c26c145e0fcef51426d64a184d0faa651777f86f.chunks.jsonl", + "bytes": 10132848, + "sha256": "4ff2b6c077641ede60dd0baff610c239dc85144e51ca131c12b97909ca2dfad6" + }, + { + "path": "data/corpus/v2_trace2code/tokio-rs__tokio/670a907c55c7f7b27da203208e65da60de6598b2.chunks.jsonl", + "bytes": 11872809, + "sha256": "d6e79a7b4b4a30a2469223ee35907ea53e407863e44c8a9ee2d05f1fa5bb67d7" + }, + { + "path": "data/corpus/v2_trace2code/tokio-rs__tokio/873cb8ae2fc291eaffbd71e3c83d17b2f0ed7abf.chunks.jsonl", + "bytes": 11492429, + "sha256": "4682b9f757c7964f4fea91b4c2808fcd912a90b289244bbd059c69c4bbf7da58" + }, + { + "path": "data/corpus/v2_trace2code/tokio-rs__tokio/911ab21d7012a50e53971ad1292a9f18de22d4c8.chunks.jsonl", + "bytes": 11825780, + "sha256": "5a0b0c7e5bc8cd1ea37f498e590bef57944c81c6d016809491b8d172228210f9" + }, + { + "path": "data/corpus/v2_trace2code/tokio-rs__tokio/c79121391db8f8d36d4213feeb25381caee110c7.chunks.jsonl", + "bytes": 12939829, + "sha256": "1b51f6341dcdf15a1eb944bea2e92c6ab059b785e091379dc5cc562c08e975fb" + }, + { + "path": "data/corpus/v2_trace2code/tokio-rs__tokio/de6ef21a81fc451f904bcad8ba85dd8c96b17d76.chunks.jsonl", + "bytes": 11946371, + "sha256": "6ef60166b8a7ada64921b749740e1d6dec2f0e129daddefc02890b887443a93d" + }, + { + "path": "data/reports/v2_trace2code/README.md", + "bytes": 341, + "sha256": "5862000dd8b1d0aa9461bcb82f990931c632fc4618e586751782708ad8fdf3f2" + }, + { + "path": "data/reports/v2_trace2code/release_manifest.json", + "bytes": 1006, + "sha256": "2a3e8b636f76da01e8ea0ec04c3324d79762f541137100e3483503d52a0cd4d2" + } +] diff --git a/blueprints/convergence-practice/arb-trace2code/metrics.json b/blueprints/convergence-practice/arb-trace2code/metrics.json new file mode 100644 index 000000000..d4dfd20b3 --- /dev/null +++ b/blueprints/convergence-practice/arb-trace2code/metrics.json @@ -0,0 +1,156 @@ +{ + "definitions": { + "MRR": "For each sample, reciprocal of the first exact gold-file rank in the full deduplicated file ranking, not MRR@20; then macro-average.", + "Recall@k": "For each sample, number of distinct gold file paths ranked at most k divided by its number of distinct gold paths; then macro-average. This is not hit@k.", + "aggregation": "Arithmetic mean over the same101 samples, not a mean of repository means. No weighting or significance test.", + "file_ranking": "Native chunk scores sorted by score, path and chunk identifier; file ranks deduplicate paths in that order." + }, + "denominator": 101, + "gold_file_count_distribution": { + "1": 80, + "2": 16, + "3": 5 + }, + "gold_file_references": 127, + "macro_over_samples": { + "bm25": { + "MRR": 0.16384832699974178, + "Recall@10": 0.3217821782178218, + "Recall@20": 0.49339933993399343, + "Recall@5": 0.22277227722772278 + }, + "lexical": { + "MRR": 0.20745277382862923, + "Recall@10": 0.48184818481848185, + "Recall@20": 0.6963696369636964, + "Recall@5": 0.3432343234323432 + } + }, + "paired_mean_delta_lexical_minus_bm25": { + "MRR": 0.043604446828887436, + "Recall@10": 0.1600660066006601, + "Recall@20": 0.20297029702970298, + "Recall@5": 0.12046204620462046 + }, + "per_repository": { + "caddyserver/caddy": { + "bm25": { + "MRR": 0.1612793102589021, + "Recall@10": 0.5, + "Recall@20": 0.6428571428571429, + "Recall@5": 0.5 + }, + "lexical": { + "MRR": 0.21879890296039364, + "Recall@10": 0.35714285714285715, + "Recall@20": 0.6428571428571429, + "Recall@5": 0.2857142857142857 + }, + "samples": 7 + }, + "clap-rs/clap": { + "bm25": { + "MRR": 0.02, + "Recall@10": 0.0, + "Recall@20": 0.0, + "Recall@5": 0.0 + }, + "lexical": { + "MRR": 0.017543859649122806, + "Recall@10": 0.0, + "Recall@20": 0.0, + "Recall@5": 0.0 + }, + "samples": 1 + }, + "etcd-io/etcd": { + "bm25": { + "MRR": 0.08517316017316018, + "Recall@10": 0.125, + "Recall@20": 0.25, + "Recall@5": 0.125 + }, + "lexical": { + "MRR": 0.10606060606060606, + "Recall@10": 0.25, + "Recall@20": 0.625, + "Recall@5": 0.25 + }, + "samples": 4 + }, + "gin-gonic/gin": { + "bm25": { + "MRR": 0.18723910722685358, + "Recall@10": 0.3244047619047619, + "Recall@20": 0.5089285714285714, + "Recall@5": 0.19642857142857142 + }, + "lexical": { + "MRR": 0.2631931596647597, + "Recall@10": 0.6577380952380952, + "Recall@20": 0.8422619047619048, + "Recall@5": 0.4345238095238095 + }, + "samples": 56 + }, + "pallets/click": { + "bm25": { + "MRR": 0.13022609726130616, + "Recall@10": 0.30128205128205127, + "Recall@20": 0.4551282051282052, + "Recall@5": 0.19230769230769232 + }, + "lexical": { + "MRR": 0.10955076525868755, + "Recall@10": 0.22435897435897434, + "Recall@20": 0.5064102564102564, + "Recall@5": 0.18589743589743588 + }, + "samples": 26 + }, + "pytest-dev/pytest": { + "bm25": { + "MRR": 0.5, + "Recall@10": 1.0, + "Recall@20": 1.0, + "Recall@5": 1.0 + }, + "lexical": { + "MRR": 0.5, + "Recall@10": 1.0, + "Recall@20": 1.0, + "Recall@5": 1.0 + }, + "samples": 1 + }, + "tokio-rs/tokio": { + "bm25": { + "MRR": 0.11462744682853378, + "Recall@10": 0.25, + "Recall@20": 0.5, + "Recall@5": 0.25 + }, + "lexical": { + "MRR": 0.14870245235413773, + "Recall@10": 0.25, + "Recall@20": 0.3333333333333333, + "Recall@5": 0.25 + }, + "samples": 6 + } + }, + "repo_sample_counts": { + "caddyserver/caddy": 7, + "clap-rs/clap": 1, + "etcd-io/etcd": 4, + "gin-gonic/gin": 56, + "pallets/click": 26, + "pytest-dev/pytest": 1, + "tokio-rs/tokio": 6 + }, + "schema_version": 1, + "skipped": { + "bm25": 0, + "lexical": 0 + } +} diff --git a/blueprints/convergence-practice/arb-trace2code/paired-samples.json b/blueprints/convergence-practice/arb-trace2code/paired-samples.json new file mode 100644 index 000000000..3d0d0a827 --- /dev/null +++ b/blueprints/convergence-practice/arb-trace2code/paired-samples.json @@ -0,0 +1,1699 @@ +{ + "samples": [ + { + "base_commit": "aed1af59763d54520a9c72f1fd0222d43904ebfd", + "gold_files": [ + "listeners.go" + ], + "gold_ranks": { + "bm25": { + "listeners.go": 3 + }, + "lexical": { + "listeners.go": 1 + } + }, + "repo": "caddyserver/caddy", + "sample_id": "42f8d46378b0c049f2eb66a8" + }, + { + "base_commit": "ca0ca67fbdb831c026d334dfd77ecc653f321876", + "gold_files": [ + "caddyconfig/caddyfile/formatter.go", + "caddyconfig/caddyfile/parse.go" + ], + "gold_ranks": { + "bm25": { + "caddyconfig/caddyfile/formatter.go": 28, + "caddyconfig/caddyfile/parse.go": 15 + }, + "lexical": { + "caddyconfig/caddyfile/formatter.go": 34, + "caddyconfig/caddyfile/parse.go": 11 + } + }, + "repo": "caddyserver/caddy", + "sample_id": "5bdeecc5409069d28e667e6d" + }, + { + "base_commit": "18ab0f955fc1075d7727c7658dbfb734c673a5c9", + "gold_files": [ + "modules/caddytls/acmeissuer.go" + ], + "gold_ranks": { + "bm25": { + "modules/caddytls/acmeissuer.go": 4 + }, + "lexical": { + "modules/caddytls/acmeissuer.go": 5 + } + }, + "repo": "caddyserver/caddy", + "sample_id": "9d7b952a6132256cd66c9612" + }, + { + "base_commit": "ef496e58ef9e7844cf8f1831030713ee5e9b354b", + "gold_files": [ + "modules/caddyhttp/caddyauth/caddyauth.go" + ], + "gold_ranks": { + "bm25": { + "modules/caddyhttp/caddyauth/caddyauth.go": 117 + }, + "lexical": { + "modules/caddyhttp/caddyauth/caddyauth.go": 69 + } + }, + "repo": "caddyserver/caddy", + "sample_id": "d4149b3f568dfe275250b62e" + }, + { + "base_commit": "4f504588669e28373f455b694539475ffe4d2926", + "gold_files": [ + "caddyconfig/httpcaddyfile/builtins.go", + "caddyconfig/httpcaddyfile/tlsapp.go" + ], + "gold_ranks": { + "bm25": { + "caddyconfig/httpcaddyfile/builtins.go": 49, + "caddyconfig/httpcaddyfile/tlsapp.go": 78 + }, + "lexical": { + "caddyconfig/httpcaddyfile/builtins.go": 70, + "caddyconfig/httpcaddyfile/tlsapp.go": 60 + } + }, + "repo": "caddyserver/caddy", + "sample_id": "de059c1c2e3ffc63820ce0e4" + }, + { + "base_commit": "7586e68e273fbe84b588e73c014d15d815413648", + "gold_files": [ + "caddyconfig/caddyfile/parse.go" + ], + "gold_ranks": { + "bm25": { + "caddyconfig/caddyfile/parse.go": 5 + }, + "lexical": { + "caddyconfig/caddyfile/parse.go": 15 + } + }, + "repo": "caddyserver/caddy", + "sample_id": "eb750774632943e952f09d40" + }, + { + "base_commit": "95941a71e87d6ccd6318bfd95c56929b01edfafd", + "gold_files": [ + "modules/caddyhttp/metrics.go", + "modules/caddyhttp/routes.go" + ], + "gold_ranks": { + "bm25": { + "modules/caddyhttp/metrics.go": 4, + "modules/caddyhttp/routes.go": 16 + }, + "lexical": { + "modules/caddyhttp/metrics.go": 7, + "modules/caddyhttp/routes.go": 15 + } + }, + "repo": "caddyserver/caddy", + "sample_id": "f7d79ec3fd20e651c379facb" + }, + { + "base_commit": "e82e1edf76bcbddf5fe53428d297520d76a6a300", + "gold_files": [ + "clap_builder/src/builder/arg.rs", + "clap_builder/src/parser/parser.rs" + ], + "gold_ranks": { + "bm25": { + "clap_builder/src/builder/arg.rs": 55, + "clap_builder/src/parser/parser.rs": 50 + }, + "lexical": { + "clap_builder/src/builder/arg.rs": 57, + "clap_builder/src/parser/parser.rs": 110 + } + }, + "repo": "clap-rs/clap", + "sample_id": "2d65eeb2c97d07f557b1ddb2" + }, + { + "base_commit": "be3acf4ed017efb7008beb08585d890e5364ce93", + "gold_files": [ + "cache/cache.go", + "cache/config.go" + ], + "gold_ranks": { + "bm25": { + "cache/cache.go": 5, + "cache/config.go": 28 + }, + "lexical": { + "cache/cache.go": 5, + "cache/config.go": 4 + } + }, + "repo": "etcd-io/etcd", + "sample_id": "19c290f5b37ff6d5a93db942" + }, + { + "base_commit": "c64f8fd68c2e75cc65b88bc0d956547a5f77cab1", + "gold_files": [ + "cache/cache.go", + "cache/demux.go" + ], + "gold_ranks": { + "bm25": { + "cache/cache.go": 27, + "cache/demux.go": 14 + }, + "lexical": { + "cache/cache.go": 29, + "cache/demux.go": 14 + } + }, + "repo": "etcd-io/etcd", + "sample_id": "a9432b3d44a06f48f40060bb" + }, + { + "base_commit": "b83b67e483936a94e59a61198934e88bf0ef95b1", + "gold_files": [ + "server/etcdserver/util.go", + "server/etcdserver/v3_server.go", + "server/features/etcd_features.go" + ], + "gold_ranks": { + "bm25": { + "server/etcdserver/util.go": 366, + "server/etcdserver/v3_server.go": 191, + "server/features/etcd_features.go": 42 + }, + "lexical": { + "server/etcdserver/util.go": 285, + "server/etcdserver/v3_server.go": 84, + "server/features/etcd_features.go": 647 + } + }, + "repo": "etcd-io/etcd", + "sample_id": "b389823a6435d1ada9d7d2b9" + }, + { + "base_commit": "8687c97fa5fc7814bf2deaf4b13e666a1b9a0dda", + "gold_files": [ + "server/proxy/grpcproxy/adapter/chan_stream.go" + ], + "gold_ranks": { + "bm25": { + "server/proxy/grpcproxy/adapter/chan_stream.go": 22 + }, + "lexical": { + "server/proxy/grpcproxy/adapter/chan_stream.go": 11 + } + }, + "repo": "etcd-io/etcd", + "sample_id": "d63c68165982a7371fcff2e0" + }, + { + "base_commit": "688a429d19d8c804447bb889d3635e2c31a5564d", + "gold_files": [ + "logger.go" + ], + "gold_ranks": { + "bm25": { + "logger.go": 1 + }, + "lexical": { + "logger.go": 3 + } + }, + "repo": "gin-gonic/gin", + "sample_id": "035c3b77d35f10df09605eff" + }, + { + "base_commit": "fc1c43298de675e5252d0b44f97dc5e204bd4e1e", + "gold_files": [ + "gin.go" + ], + "gold_ranks": { + "bm25": { + "gin.go": 16 + }, + "lexical": { + "gin.go": 7 + } + }, + "repo": "gin-gonic/gin", + "sample_id": "05041faae6e19b6882e3074a" + }, + { + "base_commit": "28e57f58b184b2305ace192e02496bb89f6fd8cb", + "gold_files": [ + "binding/form_mapping.go" + ], + "gold_ranks": { + "bm25": { + "binding/form_mapping.go": 26 + }, + "lexical": { + "binding/form_mapping.go": 5 + } + }, + "repo": "gin-gonic/gin", + "sample_id": "0765b7d2a08dd5736ac91857" + }, + { + "base_commit": "5e5ff3ace496a31b138b0820136a146bfb5de0ef", + "gold_files": [ + "context.go" + ], + "gold_ranks": { + "bm25": { + "context.go": 48 + }, + "lexical": { + "context.go": 27 + } + }, + "repo": "gin-gonic/gin", + "sample_id": "0f4d4723a06149a5172fed0a" + }, + { + "base_commit": "2a794cd0b0faa7d829291375b27a3467ea972b0d", + "gold_files": [ + "debug.go" + ], + "gold_ranks": { + "bm25": { + "debug.go": 17 + }, + "lexical": { + "debug.go": 20 + } + }, + "repo": "gin-gonic/gin", + "sample_id": "0f88458fc1fe4acce078ce3f" + }, + { + "base_commit": "9f598a31aafb92d675f38f1c8371e4ac76f858bf", + "gold_files": [ + "binding/form_mapping.go" + ], + "gold_ranks": { + "bm25": { + "binding/form_mapping.go": 50 + }, + "lexical": { + "binding/form_mapping.go": 6 + } + }, + "repo": "gin-gonic/gin", + "sample_id": "12951fd6d477e49ce706a14b" + }, + { + "base_commit": "db309081bc5c137b2aa15701ef53f7f19788da25", + "gold_files": [ + "render/data.go" + ], + "gold_ranks": { + "bm25": { + "render/data.go": 46 + }, + "lexical": { + "render/data.go": 8 + } + }, + "repo": "gin-gonic/gin", + "sample_id": "13dc98681c5ab25210a04f43" + }, + { + "base_commit": "dab5944a7bca8ae37d947dda02ac591afc1177d3", + "gold_files": [ + "recovery.go" + ], + "gold_ranks": { + "bm25": { + "recovery.go": 28 + }, + "lexical": { + "recovery.go": 13 + } + }, + "repo": "gin-gonic/gin", + "sample_id": "1e7cac8cfdead734e7e51ad6" + }, + { + "base_commit": "ecb3f7b5e2f3915bf1db240ed5eee572f8dbea36", + "gold_files": [ + "gin.go", + "path.go" + ], + "gold_ranks": { + "bm25": { + "gin.go": 10, + "path.go": 42 + }, + "lexical": { + "gin.go": 14, + "path.go": 76 + } + }, + "repo": "gin-gonic/gin", + "sample_id": "2097c0a5e982d33870da7ffa" + }, + { + "base_commit": "16cd8cdd4ef9257b1e86b5119df1536de71c6a87", + "gold_files": [ + "binding/form_mapping.go" + ], + "gold_ranks": { + "bm25": { + "binding/form_mapping.go": 46 + }, + "lexical": { + "binding/form_mapping.go": 8 + } + }, + "repo": "gin-gonic/gin", + "sample_id": "24b8f50a23a0e2ebcf4b7067" + }, + { + "base_commit": "a481ee2897af1e368de5c919fbeb21b89aa26fc7", + "gold_files": [ + "gin.go" + ], + "gold_ranks": { + "bm25": { + "gin.go": 6 + }, + "lexical": { + "gin.go": 6 + } + }, + "repo": "gin-gonic/gin", + "sample_id": "2a108bc05f87154f903de451" + }, + { + "base_commit": "bdc1ad7987b6931801af9d50e2df25667fdfaaf4", + "gold_files": [ + "response_writer.go" + ], + "gold_ranks": { + "bm25": { + "response_writer.go": 10 + }, + "lexical": { + "response_writer.go": 18 + } + }, + "repo": "gin-gonic/gin", + "sample_id": "30fb9ae30b972db986c4365d" + }, + { + "base_commit": "334160bab772f6f93767b870f9d07c176cd4aa2b", + "gold_files": [ + "gin.go", + "tree.go" + ], + "gold_ranks": { + "bm25": { + "gin.go": 9, + "tree.go": 12 + }, + "lexical": { + "gin.go": 10, + "tree.go": 17 + } + }, + "repo": "gin-gonic/gin", + "sample_id": "34b3f2f207dc4ebd77268875" + }, + { + "base_commit": "b2b489dbf4826c2c630717a77fd5e42774625410", + "gold_files": [ + "binding/form_mapping.go" + ], + "gold_ranks": { + "bm25": { + "binding/form_mapping.go": 16 + }, + "lexical": { + "binding/form_mapping.go": 3 + } + }, + "repo": "gin-gonic/gin", + "sample_id": "3bd1eecf0ebcd9c6b334fa92" + }, + { + "base_commit": "38534e2bf98a06e1f62d6b24384e90b5f78699bf", + "gold_files": [ + "debug.go" + ], + "gold_ranks": { + "bm25": { + "debug.go": 2 + }, + "lexical": { + "debug.go": 2 + } + }, + "repo": "gin-gonic/gin", + "sample_id": "4480edb6cb7cbebd6c6f7126" + }, + { + "base_commit": "a64286a7760be2031209686ce4d36e99d42dd419", + "gold_files": [ + "logger.go" + ], + "gold_ranks": { + "bm25": { + "logger.go": 59 + }, + "lexical": { + "logger.go": 19 + } + }, + "repo": "gin-gonic/gin", + "sample_id": "4543cc1c282d6fe603fdc317" + }, + { + "base_commit": "fb2583442c4d9bccb75e6d26f1aa6e7c01950db6", + "gold_files": [ + "tree.go" + ], + "gold_ranks": { + "bm25": { + "tree.go": 1 + }, + "lexical": { + "tree.go": 1 + } + }, + "repo": "gin-gonic/gin", + "sample_id": "4882164fe246221e92af0952" + }, + { + "base_commit": "c358d5656d0feb8b310d4ec379bccde46ccc8cc7", + "gold_files": [ + "recovery.go" + ], + "gold_ranks": { + "bm25": { + "recovery.go": 1 + }, + "lexical": { + "recovery.go": 1 + } + }, + "repo": "gin-gonic/gin", + "sample_id": "4ac3545de34d4974f16f5838" + }, + { + "base_commit": "ac95fa6bbcf5ec102bc81fbbb70e8fccba90f0b1", + "gold_files": [ + "context.go" + ], + "gold_ranks": { + "bm25": { + "context.go": 30 + }, + "lexical": { + "context.go": 11 + } + }, + "repo": "gin-gonic/gin", + "sample_id": "4af738a83e9f45b8aec78341" + }, + { + "base_commit": "d4a64265f21993368c90602c18e778bf04ef36db", + "gold_files": [ + "context.go" + ], + "gold_ranks": { + "bm25": { + "context.go": 11 + }, + "lexical": { + "context.go": 7 + } + }, + "repo": "gin-gonic/gin", + "sample_id": "4c232faf5b66c5e8febbe45d" + }, + { + "base_commit": "3f5b0afa2ac85ea79638ca08f4140ce64b8246e5", + "gold_files": [ + "context.go" + ], + "gold_ranks": { + "bm25": { + "context.go": 30 + }, + "lexical": { + "context.go": 6 + } + }, + "repo": "gin-gonic/gin", + "sample_id": "589abc5542483246a640f904" + }, + { + "base_commit": "19b877fa50cbbb9282763099fb177a1e5cc5c850", + "gold_files": [ + "recovery.go" + ], + "gold_ranks": { + "bm25": { + "recovery.go": 2 + }, + "lexical": { + "recovery.go": 7 + } + }, + "repo": "gin-gonic/gin", + "sample_id": "5bc1aa35cdd32914d7ba6958" + }, + { + "base_commit": "2e2bd1f408fdaa5e522bc56972e3f109fd7502fd", + "gold_files": [ + "context.go" + ], + "gold_ranks": { + "bm25": { + "context.go": 3 + }, + "lexical": { + "context.go": 4 + } + }, + "repo": "gin-gonic/gin", + "sample_id": "5eeea2e7090c74efe2defcdc" + }, + { + "base_commit": "ed150e72544949a96e7eb0b7d1151cf1907068fb", + "gold_files": [ + "binding/form_mapping.go" + ], + "gold_ranks": { + "bm25": { + "binding/form_mapping.go": 20 + }, + "lexical": { + "binding/form_mapping.go": 2 + } + }, + "repo": "gin-gonic/gin", + "sample_id": "5fe7a88d29444ad19bb646f6" + }, + { + "base_commit": "ba093d19477b896ac89a7fc3246af23d290b8e26", + "gold_files": [ + "logger.go" + ], + "gold_ranks": { + "bm25": { + "logger.go": 40 + }, + "lexical": { + "logger.go": 6 + } + }, + "repo": "gin-gonic/gin", + "sample_id": "6c6d138c6c8afd2d8b73687b" + }, + { + "base_commit": "3cb30679b5e3021db16c776ed7e70d380586e9ce", + "gold_files": [ + "binding/form_mapping.go" + ], + "gold_ranks": { + "bm25": { + "binding/form_mapping.go": 43 + }, + "lexical": { + "binding/form_mapping.go": 4 + } + }, + "repo": "gin-gonic/gin", + "sample_id": "73e85c15eb230b9bb3333088" + }, + { + "base_commit": "28e57f58b184b2305ace192e02496bb89f6fd8cb", + "gold_files": [ + "gin.go", + "ginS/gins.go", + "render/html.go" + ], + "gold_ranks": { + "bm25": { + "gin.go": 5, + "ginS/gins.go": 1, + "render/html.go": 32 + }, + "lexical": { + "gin.go": 2, + "ginS/gins.go": 3, + "render/html.go": 37 + } + }, + "repo": "gin-gonic/gin", + "sample_id": "7db765ce2d8ea9b3f68029fd" + }, + { + "base_commit": "2e22e5085960205fbb11c25776f6ea76b8053253", + "gold_files": [ + "gin.go" + ], + "gold_ranks": { + "bm25": { + "gin.go": 3 + }, + "lexical": { + "gin.go": 2 + } + }, + "repo": "gin-gonic/gin", + "sample_id": "81012a1000bf705dfb9e03a7" + }, + { + "base_commit": "4f339e6a35b163d31b30916b37f4176d385f41bd", + "gold_files": [ + "context.go" + ], + "gold_ranks": { + "bm25": { + "context.go": 3 + }, + "lexical": { + "context.go": 2 + } + }, + "repo": "gin-gonic/gin", + "sample_id": "82a5db6d32bae53d6f85a5a3" + }, + { + "base_commit": "76dd08d512504b80ef76a76c9e6bd1831e121b71", + "gold_files": [ + "context.go" + ], + "gold_ranks": { + "bm25": { + "context.go": 14 + }, + "lexical": { + "context.go": 12 + } + }, + "repo": "gin-gonic/gin", + "sample_id": "85a691c247e4c1686ce67107" + }, + { + "base_commit": "c221133ee80c46e3a6c50717ca6f1b41d4ab7711", + "gold_files": [ + "context.go" + ], + "gold_ranks": { + "bm25": { + "context.go": 56 + }, + "lexical": { + "context.go": 53 + } + }, + "repo": "gin-gonic/gin", + "sample_id": "899589599e1f682b606c97af" + }, + { + "base_commit": "e46bd521859fdfc83c508f1d42c92cb7f91e9fcb", + "gold_files": [ + "binding/form_mapping.go" + ], + "gold_ranks": { + "bm25": { + "binding/form_mapping.go": 31 + }, + "lexical": { + "binding/form_mapping.go": 4 + } + }, + "repo": "gin-gonic/gin", + "sample_id": "8c235537a265c9b4acb1c0e8" + }, + { + "base_commit": "0cbb30aa940a643e31874f8ad8e355219d306756", + "gold_files": [ + "context.go", + "gin.go" + ], + "gold_ranks": { + "bm25": { + "context.go": 58, + "gin.go": 42 + }, + "lexical": { + "context.go": 37, + "gin.go": 19 + } + }, + "repo": "gin-gonic/gin", + "sample_id": "9151d23c9630aed754b3355b" + }, + { + "base_commit": "9c081de9cdd1948f521d47d170d18cbc2981c33a", + "gold_files": [ + "gin.go" + ], + "gold_ranks": { + "bm25": { + "gin.go": 10 + }, + "lexical": { + "gin.go": 2 + } + }, + "repo": "gin-gonic/gin", + "sample_id": "94882daee852885df4f04f6e" + }, + { + "base_commit": "44d0dd70924dd154e3b98bc340accc53484efa9c", + "gold_files": [ + "tree.go" + ], + "gold_ranks": { + "bm25": { + "tree.go": 11 + }, + "lexical": { + "tree.go": 4 + } + }, + "repo": "gin-gonic/gin", + "sample_id": "95131f4a1d5c1094e4b4a013" + }, + { + "base_commit": "61c2b1c28f0c5a754330545e31f02cd6d6f7944e", + "gold_files": [ + "recovery.go" + ], + "gold_ranks": { + "bm25": { + "recovery.go": 9 + }, + "lexical": { + "recovery.go": 11 + } + }, + "repo": "gin-gonic/gin", + "sample_id": "98ca5fc67c8dd81a8f9ab021" + }, + { + "base_commit": "67c9d4ee110e9adfe33063ef847dba56717c148a", + "gold_files": [ + "errors.go" + ], + "gold_ranks": { + "bm25": { + "errors.go": 22 + }, + "lexical": { + "errors.go": 17 + } + }, + "repo": "gin-gonic/gin", + "sample_id": "9e63d7a07d201c84d7d18546" + }, + { + "base_commit": "3ab698dc5110af1977d57226e4995c57dd34c233", + "gold_files": [ + "context.go" + ], + "gold_ranks": { + "bm25": { + "context.go": 31 + }, + "lexical": { + "context.go": 3 + } + }, + "repo": "gin-gonic/gin", + "sample_id": "a69837e52fea4b05cbedcb9f" + }, + { + "base_commit": "915e4c90d28ec4cffc6eb146e208ab5a65eac772", + "gold_files": [ + "context.go" + ], + "gold_ranks": { + "bm25": { + "context.go": 38 + }, + "lexical": { + "context.go": 3 + } + }, + "repo": "gin-gonic/gin", + "sample_id": "a72e294e63bff7a51eb84669" + }, + { + "base_commit": "472d086af2acd924cb4b9d7be0525f7d790f69bc", + "gold_files": [ + "context.go", + "render/render.go" + ], + "gold_ranks": { + "bm25": { + "context.go": 41, + "render/render.go": 116 + }, + "lexical": { + "context.go": 24, + "render/render.go": 117 + } + }, + "repo": "gin-gonic/gin", + "sample_id": "a9c4956e69fe0ab051fc39a7" + }, + { + "base_commit": "ecdbbbe9483dd12222f2085f717a2c7cb5ac55fe", + "gold_files": [ + "binding/binding.go", + "binding/binding_nomsgpack.go", + "render/yaml.go" + ], + "gold_ranks": { + "bm25": { + "binding/binding.go": 53, + "binding/binding_nomsgpack.go": 52, + "render/yaml.go": 12 + }, + "lexical": { + "binding/binding.go": 58, + "binding/binding_nomsgpack.go": 52, + "render/yaml.go": 3 + } + }, + "repo": "gin-gonic/gin", + "sample_id": "a9f6b8adbd0f71a58c99835e" + }, + { + "base_commit": "b4f66e965ba9d60257e0de4c25d4ad4bd6115927", + "gold_files": [ + "binding/form_mapping.go" + ], + "gold_ranks": { + "bm25": { + "binding/form_mapping.go": 74 + }, + "lexical": { + "binding/form_mapping.go": 42 + } + }, + "repo": "gin-gonic/gin", + "sample_id": "aac6e0174bcf8f6defe4c774" + }, + { + "base_commit": "e0d46ded6cb6974d55a255ab122d1aa6ca0cd60e", + "gold_files": [ + "binding/form_mapping.go" + ], + "gold_ranks": { + "bm25": { + "binding/form_mapping.go": 48 + }, + "lexical": { + "binding/form_mapping.go": 4 + } + }, + "repo": "gin-gonic/gin", + "sample_id": "b8c820bb7d5f8b2e19f4e7f6" + }, + { + "base_commit": "c4287b1300363cb3dc2c8408299d3ba6deded485", + "gold_files": [ + "context.go", + "logger.go" + ], + "gold_ranks": { + "bm25": { + "context.go": 4, + "logger.go": 20 + }, + "lexical": { + "context.go": 3, + "logger.go": 48 + } + }, + "repo": "gin-gonic/gin", + "sample_id": "c06816923c65bcd86bfd1c2a" + }, + { + "base_commit": "05caa5c00e552a27c3bfb659b99c0d79b81fafe4", + "gold_files": [ + "binding/default_validator.go" + ], + "gold_ranks": { + "bm25": { + "binding/default_validator.go": 2 + }, + "lexical": { + "binding/default_validator.go": 2 + } + }, + "repo": "gin-gonic/gin", + "sample_id": "c4a057d97f889140494823d0" + }, + { + "base_commit": "64ead9e6bd924d431f4dd612349bc5e13300e6fc", + "gold_files": [ + "binding/form_mapping.go" + ], + "gold_ranks": { + "bm25": { + "binding/form_mapping.go": 35 + }, + "lexical": { + "binding/form_mapping.go": 5 + } + }, + "repo": "gin-gonic/gin", + "sample_id": "c7ce74f40b6308a8504688cc" + }, + { + "base_commit": "c2ba8f19ec19914b73290c53a32de479cd463555", + "gold_files": [ + "logger.go" + ], + "gold_ranks": { + "bm25": { + "logger.go": 30 + }, + "lexical": { + "logger.go": 6 + } + }, + "repo": "gin-gonic/gin", + "sample_id": "cc9de6b29530ddfb3135f0ea" + }, + { + "base_commit": "fad706f1216e6d12bdd51d28d5a40ec27e6c6453", + "gold_files": [ + "binding/form_mapping.go" + ], + "gold_ranks": { + "bm25": { + "binding/form_mapping.go": 15 + }, + "lexical": { + "binding/form_mapping.go": 2 + } + }, + "repo": "gin-gonic/gin", + "sample_id": "cd5fdafbfa862ca7662667c4" + }, + { + "base_commit": "90cf4602698dcbce18df3165b2d24e2940670a41", + "gold_files": [ + "binding/form_mapping.go" + ], + "gold_ranks": { + "bm25": { + "binding/form_mapping.go": 49 + }, + "lexical": { + "binding/form_mapping.go": 24 + } + }, + "repo": "gin-gonic/gin", + "sample_id": "d0551ed5a2d63c99e3bbfcd5" + }, + { + "base_commit": "79a61b90324586d3c3a5859f8755cae2d1c46f2d", + "gold_files": [ + "tree.go" + ], + "gold_ranks": { + "bm25": { + "tree.go": 11 + }, + "lexical": { + "tree.go": 3 + } + }, + "repo": "gin-gonic/gin", + "sample_id": "d3fe4d1559febe2ac282cced" + }, + { + "base_commit": "7a865dcf1dbe6ec52e074b1ddce830d278eb72cf", + "gold_files": [ + "binding/binding.go", + "binding/binding_nomsgpack.go", + "context.go" + ], + "gold_ranks": { + "bm25": { + "binding/binding.go": 8, + "binding/binding_nomsgpack.go": 7, + "context.go": 2 + }, + "lexical": { + "binding/binding.go": 24, + "binding/binding_nomsgpack.go": 14, + "context.go": 2 + } + }, + "repo": "gin-gonic/gin", + "sample_id": "d5a6bc19ef154663f7e591ad" + }, + { + "base_commit": "d4a64265f21993368c90602c18e778bf04ef36db", + "gold_files": [ + "binding/form_mapping.go" + ], + "gold_ranks": { + "bm25": { + "binding/form_mapping.go": 25 + }, + "lexical": { + "binding/form_mapping.go": 6 + } + }, + "repo": "gin-gonic/gin", + "sample_id": "d65a03b1febee94c93435793" + }, + { + "base_commit": "b1c1e7b572f76071fb0e0e7884a0697e0458aa7c", + "gold_files": [ + "fs.go", + "routergroup.go" + ], + "gold_ranks": { + "bm25": { + "fs.go": 1, + "routergroup.go": 6 + }, + "lexical": { + "fs.go": 1, + "routergroup.go": 17 + } + }, + "repo": "gin-gonic/gin", + "sample_id": "ddc4c4966436317b69c8484b" + }, + { + "base_commit": "a4ac275e079d46d493965491d686bfe72d121e85", + "gold_files": [ + "context.go" + ], + "gold_ranks": { + "bm25": { + "context.go": 54 + }, + "lexical": { + "context.go": 51 + } + }, + "repo": "gin-gonic/gin", + "sample_id": "e843a55e790138194506b76a" + }, + { + "base_commit": "f75144a356e57c95bd21a048f0a40492dcdb33c5", + "gold_files": [ + "context.go" + ], + "gold_ranks": { + "bm25": { + "context.go": 39 + }, + "lexical": { + "context.go": 2 + } + }, + "repo": "gin-gonic/gin", + "sample_id": "e8f3298d52397c17e0fa1bc9" + }, + { + "base_commit": "86ff4a64c7efe1a1c875529835eeef9e15de1e86", + "gold_files": [ + "gin.go" + ], + "gold_ranks": { + "bm25": { + "gin.go": 8 + }, + "lexical": { + "gin.go": 9 + } + }, + "repo": "gin-gonic/gin", + "sample_id": "fa19b2ec3770df1f1285f072" + }, + { + "base_commit": "011b9f9d190c71310264e6c54bae6259f5e38a9f", + "gold_files": [ + "src/click/core.py" + ], + "gold_ranks": { + "bm25": { + "src/click/core.py": 23 + }, + "lexical": { + "src/click/core.py": 18 + } + }, + "repo": "pallets/click", + "sample_id": "08d24e63e480c9447082ccfb" + }, + { + "base_commit": "82a83fda1fafd4a43c67755140555bfd5afbd592", + "gold_files": [ + "src/click/core.py", + "src/click/types.py" + ], + "gold_ranks": { + "bm25": { + "src/click/core.py": 8, + "src/click/types.py": 22 + }, + "lexical": { + "src/click/core.py": 2, + "src/click/types.py": 21 + } + }, + "repo": "pallets/click", + "sample_id": "27801f382096eabc8e450072" + }, + { + "base_commit": "422cb2064eed146dbe03ba3aed5be35daeb70414", + "gold_files": [ + "src/click/core.py" + ], + "gold_ranks": { + "bm25": { + "src/click/core.py": 49 + }, + "lexical": { + "src/click/core.py": 42 + } + }, + "repo": "pallets/click", + "sample_id": "2bf38b97ba88b8246c44e306" + }, + { + "base_commit": "2ed395b0b5ac4d56553ff715335f456f812cdc78", + "gold_files": [ + "src/click/core.py" + ], + "gold_ranks": { + "bm25": { + "src/click/core.py": 10 + }, + "lexical": { + "src/click/core.py": 37 + } + }, + "repo": "pallets/click", + "sample_id": "2c98330b6efb89289f963c2e" + }, + { + "base_commit": "fcd85032cff78aa536a6d2b455fb83bfcc02b228", + "gold_files": [ + "src/click/core.py" + ], + "gold_ranks": { + "bm25": { + "src/click/core.py": 36 + }, + "lexical": { + "src/click/core.py": 8 + } + }, + "repo": "pallets/click", + "sample_id": "399f735b329d5c11913ac4ab" + }, + { + "base_commit": "b625b9f39082f366dd85ad13024e43c8f5e78b87", + "gold_files": [ + "src/click/termui.py" + ], + "gold_ranks": { + "bm25": { + "src/click/termui.py": 13 + }, + "lexical": { + "src/click/termui.py": 5 + } + }, + "repo": "pallets/click", + "sample_id": "461a1aafd27445f29ef24020" + }, + { + "base_commit": "7f7bbe4569ea68e8dabee232eade069ef3310aea", + "gold_files": [ + "src/click/core.py" + ], + "gold_ranks": { + "bm25": { + "src/click/core.py": 5 + }, + "lexical": { + "src/click/core.py": 19 + } + }, + "repo": "pallets/click", + "sample_id": "49589b875553b5fd2bdb1687" + }, + { + "base_commit": "a8c05427602d12842c53ce48b03b5d6b5c0c6693", + "gold_files": [ + "src/click/core.py" + ], + "gold_ranks": { + "bm25": { + "src/click/core.py": 26 + }, + "lexical": { + "src/click/core.py": 23 + } + }, + "repo": "pallets/click", + "sample_id": "4dfab03e69e4615b806dbab1" + }, + { + "base_commit": "380008389b32ca2b7de7c73f670a9eee1562004e", + "gold_files": [ + "src/click/core.py" + ], + "gold_ranks": { + "bm25": { + "src/click/core.py": 46 + }, + "lexical": { + "src/click/core.py": 55 + } + }, + "repo": "pallets/click", + "sample_id": "51fc5202e13c950295302355" + }, + { + "base_commit": "08ecfd0e7c238a2444886d239de6d0e9d83530ec", + "gold_files": [ + "src/click/core.py" + ], + "gold_ranks": { + "bm25": { + "src/click/core.py": 25 + }, + "lexical": { + "src/click/core.py": 19 + } + }, + "repo": "pallets/click", + "sample_id": "6d3773a91bd26a4977aa45af" + }, + { + "base_commit": "a1235aacb1be55dc66ddcfefbf64dec44b6ab54d", + "gold_files": [ + "src/click/core.py" + ], + "gold_ranks": { + "bm25": { + "src/click/core.py": 23 + }, + "lexical": { + "src/click/core.py": 23 + } + }, + "repo": "pallets/click", + "sample_id": "8882930f15d1cd0966e4ff81" + }, + { + "base_commit": "0bf333e1b77b4ae0e85d387531b38930fc657f12", + "gold_files": [ + "src/click/__init__.py", + "src/click/core.py", + "src/click/decorators.py" + ], + "gold_ranks": { + "bm25": { + "src/click/__init__.py": 57, + "src/click/core.py": 8, + "src/click/decorators.py": 34 + }, + "lexical": { + "src/click/__init__.py": 78, + "src/click/core.py": 5, + "src/click/decorators.py": 18 + } + }, + "repo": "pallets/click", + "sample_id": "8fbcfd329087a4fe16649d69" + }, + { + "base_commit": "c88f333de440c5066ff563f784a6561f4713cf78", + "gold_files": [ + "src/click/core.py" + ], + "gold_ranks": { + "bm25": { + "src/click/core.py": 43 + }, + "lexical": { + "src/click/core.py": 37 + } + }, + "repo": "pallets/click", + "sample_id": "9372208a6dfa7075afb3f423" + }, + { + "base_commit": "ddede2147c9370ab8275f81db4826b6895d2e7dc", + "gold_files": [ + "src/click/termui.py" + ], + "gold_ranks": { + "bm25": { + "src/click/termui.py": 4 + }, + "lexical": { + "src/click/termui.py": 4 + } + }, + "repo": "pallets/click", + "sample_id": "9f0d0d1bb836481d62b838aa" + }, + { + "base_commit": "60cad4ad6583bf885c62d6f4597f3555d1ee1a0f", + "gold_files": [ + "src/click/core.py" + ], + "gold_ranks": { + "bm25": { + "src/click/core.py": 5 + }, + "lexical": { + "src/click/core.py": 13 + } + }, + "repo": "pallets/click", + "sample_id": "a4218c7e484b796962f32982" + }, + { + "base_commit": "fa7f035ab58dc4d9b8818b92fbc6a8f3c3234af6", + "gold_files": [ + "src/click/core.py" + ], + "gold_ranks": { + "bm25": { + "src/click/core.py": 11 + }, + "lexical": { + "src/click/core.py": 12 + } + }, + "repo": "pallets/click", + "sample_id": "a94f5208796568631ef01c1e" + }, + { + "base_commit": "02046e7a19480f85fff7e4577486518abe47e401", + "gold_files": [ + "src/click/types.py" + ], + "gold_ranks": { + "bm25": { + "src/click/types.py": 1 + }, + "lexical": { + "src/click/types.py": 18 + } + }, + "repo": "pallets/click", + "sample_id": "af8170e7fe43a980be55e40d" + }, + { + "base_commit": "4271fe283dc9365563aebb369ada8d20eee015a8", + "gold_files": [ + "src/click/core.py", + "src/click/exceptions.py" + ], + "gold_ranks": { + "bm25": { + "src/click/core.py": 43, + "src/click/exceptions.py": 34 + }, + "lexical": { + "src/click/core.py": 34, + "src/click/exceptions.py": 53 + } + }, + "repo": "pallets/click", + "sample_id": "bed9a23bf8a662f0e6251a42" + }, + { + "base_commit": "c7e1ba8448cbcb2cdd9c1c7f4a592e863dcc3995", + "gold_files": [ + "src/click/core.py" + ], + "gold_ranks": { + "bm25": { + "src/click/core.py": 7 + }, + "lexical": { + "src/click/core.py": 11 + } + }, + "repo": "pallets/click", + "sample_id": "bf9f481e5c9e9150f3f773b6" + }, + { + "base_commit": "bf3a945b3a220589c67d55dcb73cc1c9a9a175b3", + "gold_files": [ + "src/click/core.py" + ], + "gold_ranks": { + "bm25": { + "src/click/core.py": 28 + }, + "lexical": { + "src/click/core.py": 21 + } + }, + "repo": "pallets/click", + "sample_id": "cdca9bf2d9c3acf91f505d3a" + }, + { + "base_commit": "6c7624491911f0b18f869e5d0a989506e7b416f8", + "gold_files": [ + "src/click/core.py" + ], + "gold_ranks": { + "bm25": { + "src/click/core.py": 34 + }, + "lexical": { + "src/click/core.py": 27 + } + }, + "repo": "pallets/click", + "sample_id": "d9073750309233ee1f19776d" + }, + { + "base_commit": "bf9da4838986af3bc09a1b16974d154c61dcbe3f", + "gold_files": [ + "src/click/core.py" + ], + "gold_ranks": { + "bm25": { + "src/click/core.py": 34 + }, + "lexical": { + "src/click/core.py": 25 + } + }, + "repo": "pallets/click", + "sample_id": "dfa7089d7bf0b3dad96e2bb8" + }, + { + "base_commit": "1a4d8c1bb1e8f8e214ede7223bd2c05dc2ce006a", + "gold_files": [ + "src/click/core.py" + ], + "gold_ranks": { + "bm25": { + "src/click/core.py": 33 + }, + "lexical": { + "src/click/core.py": 43 + } + }, + "repo": "pallets/click", + "sample_id": "e50d9d504a467c632442042e" + }, + { + "base_commit": "1458800409ed12076f18451889b0857db36aa522", + "gold_files": [ + "src/click/_termui_impl.py" + ], + "gold_ranks": { + "bm25": { + "src/click/_termui_impl.py": 2 + }, + "lexical": { + "src/click/_termui_impl.py": 2 + } + }, + "repo": "pallets/click", + "sample_id": "e7f9f3fc7342903c89bfb775" + }, + { + "base_commit": "7d7183604158f064390539d83d4a19a978c6b08a", + "gold_files": [ + "src/click/core.py" + ], + "gold_ranks": { + "bm25": { + "src/click/core.py": 14 + }, + "lexical": { + "src/click/core.py": 22 + } + }, + "repo": "pallets/click", + "sample_id": "ed5f5fce7ddb0be735cf98c7" + }, + { + "base_commit": "8d3449517a97b6fd457a025bd9c3d2acfd301fa4", + "gold_files": [ + "src/click/exceptions.py" + ], + "gold_ranks": { + "bm25": { + "src/click/exceptions.py": 11 + }, + "lexical": { + "src/click/exceptions.py": 5 + } + }, + "repo": "pallets/click", + "sample_id": "f4fccd0dc58a48c09a195d6b" + }, + { + "base_commit": "c26c145e0fcef51426d64a184d0faa651777f86f", + "gold_files": [ + "src/_pytest/fixtures.py" + ], + "gold_ranks": { + "bm25": { + "src/_pytest/fixtures.py": 2 + }, + "lexical": { + "src/_pytest/fixtures.py": 2 + } + }, + "repo": "pytest-dev/pytest", + "sample_id": "c727abdbfcdfc9879df581d2" + }, + { + "base_commit": "873cb8ae2fc291eaffbd71e3c83d17b2f0ed7abf", + "gold_files": [ + "tokio/src/sync/mpsc/block.rs" + ], + "gold_ranks": { + "bm25": { + "tokio/src/sync/mpsc/block.rs": 322 + }, + "lexical": { + "tokio/src/sync/mpsc/block.rs": 178 + } + }, + "repo": "tokio-rs/tokio", + "sample_id": "1132f037ebb462f13652adff" + }, + { + "base_commit": "670a907c55c7f7b27da203208e65da60de6598b2", + "gold_files": [ + "tokio/src/sync/mpsc/chan.rs" + ], + "gold_ranks": { + "bm25": { + "tokio/src/sync/mpsc/chan.rs": 96 + }, + "lexical": { + "tokio/src/sync/mpsc/chan.rs": 33 + } + }, + "repo": "tokio-rs/tokio", + "sample_id": "118b720b3e60d4c6679b9ec0" + }, + { + "base_commit": "670a907c55c7f7b27da203208e65da60de6598b2", + "gold_files": [ + "tokio/src/sync/mpsc/bounded.rs" + ], + "gold_ranks": { + "bm25": { + "tokio/src/sync/mpsc/bounded.rs": 5 + }, + "lexical": { + "tokio/src/sync/mpsc/bounded.rs": 2 + } + }, + "repo": "tokio-rs/tokio", + "sample_id": "30a18c73cca02d3672705be7" + }, + { + "base_commit": "de6ef21a81fc451f904bcad8ba85dd8c96b17d76", + "gold_files": [ + "tokio/src/sync/mpsc/chan.rs", + "tokio/src/sync/mpsc/list.rs" + ], + "gold_ranks": { + "bm25": { + "tokio/src/sync/mpsc/chan.rs": 11, + "tokio/src/sync/mpsc/list.rs": 88 + }, + "lexical": { + "tokio/src/sync/mpsc/chan.rs": 11, + "tokio/src/sync/mpsc/list.rs": 36 + } + }, + "repo": "tokio-rs/tokio", + "sample_id": "34b14301c5676ca5e49d6c22" + }, + { + "base_commit": "c79121391db8f8d36d4213feeb25381caee110c7", + "gold_files": [ + "tokio/src/sync/batch_semaphore.rs" + ], + "gold_ranks": { + "bm25": { + "tokio/src/sync/batch_semaphore.rs": 20 + }, + "lexical": { + "tokio/src/sync/batch_semaphore.rs": 65 + } + }, + "repo": "tokio-rs/tokio", + "sample_id": "954f4ea51031f7c08c603b53" + }, + { + "base_commit": "911ab21d7012a50e53971ad1292a9f18de22d4c8", + "gold_files": [ + "tokio/src/sync/mod.rs", + "tokio/src/sync/notify.rs" + ], + "gold_ranks": { + "bm25": { + "tokio/src/sync/mod.rs": 250, + "tokio/src/sync/notify.rs": 3 + }, + "lexical": { + "tokio/src/sync/mod.rs": 403, + "tokio/src/sync/notify.rs": 4 + } + }, + "repo": "tokio-rs/tokio", + "sample_id": "b872b8ef648b8cad2d959a36" + } + ], + "schema_version": 1 +} diff --git a/blueprints/convergence-practice/arb-trace2code/receipt.json b/blueprints/convergence-practice/arb-trace2code/receipt.json new file mode 100644 index 000000000..b071a40c8 --- /dev/null +++ b/blueprints/convergence-practice/arb-trace2code/receipt.json @@ -0,0 +1,201 @@ +{ + "artifact_hashes": { + "README.md": "214f2f09a151792c6e83b6705eb1973d5efeb3c3b16be563ff7cd4cca0342e16", + "input-files.json": "2203d0f03d0ccb5e863e61df5c47d09023204fbbbf8cf51b1699a82e76cc1e25", + "metrics.json": "e2329e35dcd3ebe9fb831814264aa71bbd9edfc4172a487c79cb38971c03197a", + "paired-samples.json": "b44b3547e8db13f5acfc48ba828153f4833bd04e6afbbd9ce38e11c67dcd8286", + "source-files.json": "d87ac16f0e85d6ebdcba5adb519d3bf556253e2674c2a972258df4efa75657d2", + "upstream-bm25-summary.json": "1b2451800f0b81060486659a71dc40b447097c903299dbf9ea841019128789e9", + "upstream-lexical-summary.json": "6fe6c2746c3bc0a855ee452b2de59707730d199b6f3128c1f7dee4f0b928132b", + "upstream-validation.json": "2fdff3fa025e8d82e77ee23756d490a4579cd9d5f5796bf57ced2306dcb2d922" + }, + "checks": { + "archive_path_safety": "Every member inspected before extraction: only regular files/directories, no absolute/traversal paths; extracted with Python tarfile data filter.", + "independent_arithmetic": "All four headline per-sample metrics and macro means recomputed from exact gold ranks and agree with upstream within1e-14.", + "manifest_path_safety": "All98 chunks_path values checked relative under data/corpus and present before evaluator invocation.", + "source_git_worktree_clean": true, + "upstream_schema_validation": { + "exit_code": 0, + "invalid": 0, + "report": "upstream-validation.json", + "samples": 101 + }, + "upstream_tests": { + "command_template": "PYTHONPATH=src ${PYTEST_PYTHON} -m pytest -q tests/test_corpus_baseline.py", + "exit_code": 0, + "failed": 0, + "passed": 15, + "pytest_version": "9.1.1", + "result_text": "15 passed in 0.15s", + "scope": "Pinned upstream corpus/lexical/BM25 tests; no full-suite/model execution claimed.", + "skipped": 0, + "source_commit": "b487f3866cc13dd971819cb902517a6a50282404" + } + }, + "dataset": { + "archive_bytes": 39295446, + "archive_entries": 120, + "archive_hash_matches_publisher_checksum_and_hub_lfs_oid": true, + "archive_sha256": "19b252e8cfff42107fedc74005dbb6972f2970af33651ce0c1571546819e41c4", + "archive_url": "https://huggingface.co/datasets/eyuansu71/agent_retrieval_bench/resolve/5901e1ee3aff048290db72edf9c63bc498b79ea3/releases/v2_trace2code/agent_retrieval_bench_v2_trace2code.tar.zst", + "checksum_file_sha256": "d844cc6f1fb9f345f97a687f140d12740e96cc16d533958581a9e77621f7eb04", + "corpus_snapshots": 98, + "extracted_bytes": 300175960, + "extracted_files": 104, + "gold_file_count_distribution": { + "1": 80, + "2": 16, + "3": 5 + }, + "input_file_hashes": "input-files.json", + "license": "Benchmark metadata/reports MIT; corpus source retains each original project license. No corpus or query text is republished in this evidence.", + "release": "v2_trace2code", + "repo_sample_counts": { + "caddyserver/caddy": 7, + "clap-rs/clap": 1, + "etcd-io/etcd": 4, + "gin-gonic/gin": 56, + "pallets/click": 26, + "pytest-dev/pytest": 1, + "tokio-rs/tokio": 6 + }, + "repositories": 7, + "repository": "https://huggingface.co/datasets/eyuansu71/agent_retrieval_bench", + "revision": "5901e1ee3aff048290db72edf9c63bc498b79ea3", + "samples": 101, + "selection_reason": "Smallest complete positive release by reviewed compressed size; all101 rows selected before evaluating results.", + "snapshot_chunk_records": 356074, + "snapshot_file_records": 24883 + }, + "environment": { + "architecture": "arm64", + "environment": "Isolated stdlib-only virtual environment; no pip packages installed.", + "os": "Darwin", + "python": "3.14.7", + "runtime_dependencies": [] + }, + "initial_source_snapshot_execution": { + "commit": "07014c986f3deadb1548c62b32c0ffbe6a81465d", + "declared_version": "0.2.1", + "differences_to_release": "README, CITATION and documentation only; no evaluator or pyproject changes.", + "is_exact_release_tag_commit": false, + "metrics_identical_after_rerun": true, + "per_sample_details_byte_identical_after_rerun": true, + "rerun_reason": "Source metadata snapshot and release tag are distinct identities; primary receipt was rerun at verified release tag commit.", + "runs": { + "bm25": { + "command_wall_seconds": 10.487976332999096, + "concurrent_with": "lexical; elapsed time is not an isolated latency benchmark", + "evaluated": 101, + "exit_code": 0, + "native_details_sha256": "3c5e53999c89b48a5df55ef870484a85786e8939b91d6c9ee3a0f29e08c40b87", + "native_wall_seconds": 10.383530625000276, + "skipped": {}, + "summary_sha256": "010707ed9fa84b112b6ac3e9674c7f7acb26a2c362b17f6415bea13863ce1780" + }, + "lexical": { + "command_wall_seconds": 10.16499712500081, + "concurrent_with": "bm25; elapsed time is not an isolated latency benchmark", + "evaluated": 101, + "exit_code": 0, + "native_details_sha256": "2eb08ad2e7f35549ec823f6911bf4150c7ccddb5a037998c3a357a03d2c22cd4", + "native_wall_seconds": 10.042018500000268, + "skipped": {}, + "summary_sha256": "dcf9265d0efbbd9655eab094909a893b6cb6af6c71f5c0486036ae4ce2a91f35" + } + }, + "source_tree": "0fd0466fb3e24c40c0bdf9468872bea9c9a61fa1", + "source_tree_identical_to_exact_release": true + }, + "kind": "actual_upstream_offline_retrieval_evaluation", + "limits": [ + "Positive-only failure-trace subset; no abstention/no-gold test and no whole-benchmark generalization.", + "Repository imbalance: gin56 and click26 together account for82/101 samples.", + "These are ARB lexical/BM25 algorithms, not rg, QMD, SocratiCode or an embedding baseline.", + "No model/provider calls, active-index modifications, native client changes or retrieval default promotion.", + "Exact file-level relevance does not prove function/span correctness, repair success or task completion.", + "Native summaries retain legacy @8k fields whose baseline implementation packs len(text) characters and may include an oversized first chunk; those are not tokenizer-measured BCY, context tokens or provider savings.", + "One complete released task with no tuned parameters; no significance test or universal quality claim.", + "Primary exact-release runs are single sequential warm observations after source-identical earlier runs; no controlled latency or cold-cache comparison." + ], + "parameters": { + "bm25": { + "b": 0.75, + "chunk_text_fields": [ + "path", + "symbol", + "kind", + "text" + ], + "k1": 1.5, + "query_frequency_weight": true, + "tie_break": "path then chunk_id ascending" + }, + "candidate_filter": "all_files", + "dry_run": false, + "keep_list": null, + "lexical": { + "basename_substring_bonus": 8.0, + "chunk_text_fields": [ + "path", + "symbol", + "kind", + "text" + ], + "full_path_substring_bonus": 25.0, + "native_tool_equivalence": false, + "normalization": "Divide combined score by max(1,sqrt(number of unique chunk tokens)).", + "query_tokenizer": "Camel-case split then lowercase ASCII alphanumeric tokens.", + "symbol_substring_bonus": 5.0, + "term_weight": "(1+log(query_count))*log((N+1)/(1+df)+1) for each matching token", + "tie_break": "path then chunk_id ascending" + }, + "limit_samples": null, + "rankers": [ + "lexical", + "bm25" + ], + "samples_argument": "data/benchmark/v2_trace2code/samples.jsonl" + }, + "receipt_created_at": "2026-09-20T02:12:40+00:00", + "runs": { + "bm25": { + "command_wall_seconds": 10.085556417001499, + "concurrent_with": null, + "evaluated": 101, + "exit_code": 0, + "native_details_sha256": "3c5e53999c89b48a5df55ef870484a85786e8939b91d6c9ee3a0f29e08c40b87", + "native_wall_seconds": 9.94414224999855, + "skipped": {}, + "started_at": "2026-09-20T02:09:16+00:00", + "summary": "upstream-bm25-summary.json", + "timing_limit": "Single sequential replay after earlier source-identical warmup; not a performance benchmark." + }, + "lexical": { + "command_wall_seconds": 10.265198290999251, + "concurrent_with": null, + "evaluated": 101, + "exit_code": 0, + "native_details_sha256": "2eb08ad2e7f35549ec823f6911bf4150c7ccddb5a037998c3a357a03d2c22cd4", + "native_wall_seconds": 10.135793833000207, + "skipped": {}, + "started_at": "2026-09-20T02:09:06+00:00", + "summary": "upstream-lexical-summary.json", + "timing_limit": "Single sequential replay after earlier source-identical warmup; not a performance benchmark." + } + }, + "schema_version": 1, + "source": { + "code_license": "MIT", + "commit": "b487f3866cc13dd971819cb902517a6a50282404", + "data_license_url": "https://github.com/eyuansu62/agent-retrieval-bench/blob/b487f3866cc13dd971819cb902517a6a50282404/DATA_LICENSE.md", + "license_url": "https://github.com/eyuansu62/agent-retrieval-bench/blob/b487f3866cc13dd971819cb902517a6a50282404/LICENSE", + "modified": false, + "repository": "https://github.com/eyuansu62/agent-retrieval-bench", + "source_hashes": "source-files.json", + "source_tree": "0fd0466fb3e24c40c0bdf9468872bea9c9a61fa1", + "tag": "v0.2.1", + "version": "0.2.1" + }, + "status": "completed" +} diff --git a/blueprints/convergence-practice/arb-trace2code/source-files.json b/blueprints/convergence-practice/arb-trace2code/source-files.json new file mode 100644 index 000000000..5e74a15d1 --- /dev/null +++ b/blueprints/convergence-practice/arb-trace2code/source-files.json @@ -0,0 +1,295 @@ +{ + "files": [ + { + "bytes": 1404, + "path": "DATA_LICENSE.md", + "sha256": "c503889e22cb019a92ddc8c51e704e89c2e85f1797d5f97cc3c022dd49f1a19e" + }, + { + "bytes": 1091, + "path": "LICENSE", + "sha256": "f821747bf0011218a8d18a6030b3343384a2c68478e643e26e057732fc89aa31" + }, + { + "bytes": 9985, + "path": "README.md", + "sha256": "494b73638318f3e8958422e6ac49e6c62d13973c0e4d947c186ed3b305c3dbc4" + }, + { + "bytes": 729, + "path": "pyproject.toml", + "sha256": "05889ec96609b4d5b557ac21b9894a3d308930dbe7d8d11deb130abdc9d2e317" + }, + { + "bytes": 107, + "path": "src/agent_retrieval_bench/__init__.py", + "sha256": "910ff4aa7e090babb5f4d1f45447bf2cfdc9be4c9b493bcd449d1fcae486694c" + }, + { + "bytes": 157707, + "path": "src/agent_retrieval_bench/abstention.py", + "sha256": "810ae98aa22472ce3d5254e0c7b4dea5e322a29167e8c390f89d781d5a28f1c3" + }, + { + "bytes": 8189, + "path": "src/agent_retrieval_bench/agentic_eval.py", + "sha256": "1e15bba4ede3d4597505bfbf2b65c6034e3d82a0f7351f02449edcbcd2e7d24c" + }, + { + "bytes": 22486, + "path": "src/agent_retrieval_bench/agentic_relevance.py", + "sha256": "5afe3add3039adfa4c2141bc65d90a68c8093c70dff5e9f267b99ffa87c36766" + }, + { + "bytes": 9811, + "path": "src/agent_retrieval_bench/audit.py", + "sha256": "29fadae6688f616429d92e0e3c21703a1d0a978417969b165108e161e95a3b1e" + }, + { + "bytes": 41557, + "path": "src/agent_retrieval_bench/baseline.py", + "sha256": "5a82f9a087eb90f9ada2ad6b6047fe8b9781b013749f3d176b1cbdbe7b7f05d1" + }, + { + "bytes": 16854, + "path": "src/agent_retrieval_bench/bcy_curve.py", + "sha256": "83aa1fd4729d2ce9f452ed41ed442213d7078e0ea3535438408c509f84d482c3" + }, + { + "bytes": 21226, + "path": "src/agent_retrieval_bench/cae_validity.py", + "sha256": "190ba5e42cf919bcbe06d0807821554dafac261fe9c58f9a3e49880d142025c6" + }, + { + "bytes": 244331, + "path": "src/agent_retrieval_bench/cli.py", + "sha256": "6214b053c7799e7f654c39e871207624748588eaecc9379e57333c27b95ab58a" + }, + { + "bytes": 2070, + "path": "src/agent_retrieval_bench/clone.py", + "sha256": "fa7e418ee0e9ca658ac8e67d9d56fe9d07545c975cdb96939d3636cb50800a2c" + }, + { + "bytes": 81477, + "path": "src/agent_retrieval_bench/closed_tool_eval.py", + "sha256": "8809ed8e6808839ce5bd5ae9b4c8017ce1e2e69b9c4173bfca4744c10253e72c" + }, + { + "bytes": 20893, + "path": "src/agent_retrieval_bench/code2test_pr.py", + "sha256": "4227aa4c1cc835ec34498d74b60dfc11280d0427a1af47f256ee44fc02e3d0d8" + }, + { + "bytes": 18844, + "path": "src/agent_retrieval_bench/context_cost.py", + "sha256": "beb48b109d706457b283d8590cb74e14e353a79a1b5054f92964522b691f6d97" + }, + { + "bytes": 12621, + "path": "src/agent_retrieval_bench/corpus.py", + "sha256": "6ddde17c0654a9f9086f6054c1ffaadbed00fc3a0a4af8c71dbce57820c69f52" + }, + { + "bytes": 43218, + "path": "src/agent_retrieval_bench/crawler.py", + "sha256": "620a1b3aa021f20a3e5e2b4e367722608be62edc77aced5a29917741d24fa3b4" + }, + { + "bytes": 3550, + "path": "src/agent_retrieval_bench/curate.py", + "sha256": "d23643dcf6e61ba71f82971de70c4fbbc802f139b0e72a34a7f77468269344f1" + }, + { + "bytes": 53411, + "path": "src/agent_retrieval_bench/dataset_validity.py", + "sha256": "9e2cf99ebef0f2616af8d66fe5be751660bfae5ef42f8aa843ff471e4554816a" + }, + { + "bytes": 43868, + "path": "src/agent_retrieval_bench/derive.py", + "sha256": "9c5a541e087d2491431a016db00f18a716197126b41ebb272b9c49caa6101f9c" + }, + { + "bytes": 17826, + "path": "src/agent_retrieval_bench/diagnostics.py", + "sha256": "e3866133f65266e4f8cd9def6c661bd5a00799bd2395b833e56f845c894cb2dc" + }, + { + "bytes": 58628, + "path": "src/agent_retrieval_bench/edit2ripple.py", + "sha256": "a7028cd50ce5fe2a7fb8757e94415a0155e5ef0ffff91120e16adfb824eb90f0" + }, + { + "bytes": 45255, + "path": "src/agent_retrieval_bench/embedding_eval.py", + "sha256": "14bb46a90a9ffc4461d7f4e61cb32a6223983cfbd7dd6af8f7c3338279c940d8" + }, + { + "bytes": 7620, + "path": "src/agent_retrieval_bench/filters.py", + "sha256": "0f71e3de8887805202c3ef9703bc548f3c26c6d4feab58a3dd9237e389a0d2a7" + }, + { + "bytes": 14057, + "path": "src/agent_retrieval_bench/git_raw.py", + "sha256": "d5098f3dc0af809fbac95f6c64147c357ad43b9238bb6327eb1e7672c97aa5aa" + }, + { + "bytes": 11301, + "path": "src/agent_retrieval_bench/github_api.py", + "sha256": "4721a8f023b49c717072372ddf4a945437da36c7c732b0341f32a575d22bc62e" + }, + { + "bytes": 18842, + "path": "src/agent_retrieval_bench/granularity.py", + "sha256": "8049b95e3df237f0cca9cf93e3737de3c3abd263539eff9bced0213dc6b15661" + }, + { + "bytes": 13266, + "path": "src/agent_retrieval_bench/grep_eval.py", + "sha256": "5e7c13b20fae343be23ae8702161f895f7605c921cdfbb2329d2cbe2533d2d8a" + }, + { + "bytes": 6509, + "path": "src/agent_retrieval_bench/hardmine.py", + "sha256": "0be145c68eadde670b18b9c2336a6fb15e52a6f25e2b22439f4da3a18483e750" + }, + { + "bytes": 42926, + "path": "src/agent_retrieval_bench/hardness.py", + "sha256": "638a71acf84bc199a2d9dcdaef512e43db6660163a9ea9928411fb583726006b" + }, + { + "bytes": 2362, + "path": "src/agent_retrieval_bench/io.py", + "sha256": "e009c326e0329b6df8da8e48ff25d563f9681a2b6a71641605fea11ccdcee2d8" + }, + { + "bytes": 12895, + "path": "src/agent_retrieval_bench/layered_leaderboard.py", + "sha256": "f0724cca1592d7794007d99e7a06fd11919c4a8386d29f99f677cb8023d51bb8" + }, + { + "bytes": 8340, + "path": "src/agent_retrieval_bench/logs.py", + "sha256": "bf3eeacc758ba07773c881b8b5fd4b916eca2eac38209ebeec9f567986f90e9c" + }, + { + "bytes": 7884, + "path": "src/agent_retrieval_bench/model_report.py", + "sha256": "a89657411f0bca657403765ff166bb530798252f1db8f660b1928ee665555eb6" + }, + { + "bytes": 31246, + "path": "src/agent_retrieval_bench/openai_context_agent.py", + "sha256": "322efc67beb2611a772677b41dcb851a39a9f1dfd2cee4204f54c1bbd0727ee7" + }, + { + "bytes": 16097, + "path": "src/agent_retrieval_bench/pes_calibration.py", + "sha256": "ef2189e7a8ada936a0ab13719776314dd6e39369cdf07e7232f1ef328447b267" + }, + { + "bytes": 2163, + "path": "src/agent_retrieval_bench/progress.py", + "sha256": "4ed5b3736c7f6dd5bf03b466d4d15a03319e528f25e4d852ef4a3adf741baaa6" + }, + { + "bytes": 4561, + "path": "src/agent_retrieval_bench/quality.py", + "sha256": "ffffee008e340556a37e8bc4b78fd80689d78257af5db802113d79d64aac8728" + }, + { + "bytes": 18113, + "path": "src/agent_retrieval_bench/rank_analysis.py", + "sha256": "55c7b74e45df29b973a08ac44b21039f06d93cad506b164f5b1f4ad616d0cc19" + }, + { + "bytes": 17344, + "path": "src/agent_retrieval_bench/rank_fusion.py", + "sha256": "0c22e72db288f2f6ca109fb0ffc7815e2fb89391fb405a4d12665d9baf98f576" + }, + { + "bytes": 8980, + "path": "src/agent_retrieval_bench/release.py", + "sha256": "28487fc70742a9210b0b6cea5dbe53980453500594f5130c9c06a9dd8704927b" + }, + { + "bytes": 27741, + "path": "src/agent_retrieval_bench/repomap_eval.py", + "sha256": "8ee7477493cb273c1aae3a850d7eeff408be9abb08ee5cf105e28a092cd0a613" + }, + { + "bytes": 5644, + "path": "src/agent_retrieval_bench/seed_report.py", + "sha256": "cc3043617674bf7c55417473fec9000ca8dd242ecc1a8a45c2f2f08f6368421e" + }, + { + "bytes": 22051, + "path": "src/agent_retrieval_bench/selective_cv.py", + "sha256": "d2d1f09f5866d673afc669450a0f8890189db3bb487ae632ecf5dd3319208524" + }, + { + "bytes": 14659, + "path": "src/agent_retrieval_bench/selective_embedding_eval.py", + "sha256": "f4d74375e9819eef41061b4d576d1b215d1ee811f89e392b02df000067d9523c" + }, + { + "bytes": 19148, + "path": "src/agent_retrieval_bench/selective_eval.py", + "sha256": "e58fe00013eed4c08fbaa4c074504d2b1dae8dfcd2bc9d83cd4fd8b43258fdd5" + }, + { + "bytes": 57305, + "path": "src/agent_retrieval_bench/trace_preflight.py", + "sha256": "ee2031d4d2e0472e12c366847e71d820d2bebb7976d5fa72f40cc90bb0bed899" + }, + { + "bytes": 44748, + "path": "src/agent_retrieval_bench/trace_repro.py", + "sha256": "0e5d0d5768ed11948b3dd3849bf5ed24e9aaa60d8ff9dbd78e922d557d3e1555" + }, + { + "bytes": 13017, + "path": "src/agent_retrieval_bench/trajectory.py", + "sha256": "c5995d461df0e3c3f9ea9b8cc51768101be6545b7fe9c4c91275dec06d8bb413" + }, + { + "bytes": 9203, + "path": "src/agent_retrieval_bench/trajectory_collect.py", + "sha256": "603e37e087e0ce38ad488e119f8ef2f36055cea4508ee7c51e3bd744cc967d54" + }, + { + "bytes": 3292, + "path": "src/agent_retrieval_bench/trajectory_compare.py", + "sha256": "652cfc0a770f8c5ff98861cdc7f22a0822d052866220136a2701d6bdd24a94e6" + }, + { + "bytes": 16491, + "path": "src/agent_retrieval_bench/trajectory_release.py", + "sha256": "fd0beda09d3a31471928e6c6431f67d71ccadbceb97f6f76c2eb2d65c9f3e3c9" + }, + { + "bytes": 420538, + "path": "src/agent_retrieval_bench/v1_1.py", + "sha256": "5f98a6b72e366aa0567c994e6611cbd971837ac9d24d6ec40cd8063b3bfd1823" + }, + { + "bytes": 27608, + "path": "src/agent_retrieval_bench/v1_2.py", + "sha256": "84fb40b5d9acb8a3aecf124242b232c1d0e3583683dee2a2be280bae357a0986" + }, + { + "bytes": 29887, + "path": "src/agent_retrieval_bench/v1_3.py", + "sha256": "93cdfe8f6ecf7da87d5a461961330ea26d398ae90c98389afa2572fc94d51078" + }, + { + "bytes": 13629, + "path": "src/agent_retrieval_bench/v2_positive_report.py", + "sha256": "0b76a68b323c478668ba0bd7055bcb5614ae250cf809f4b145b759051336dcf6" + } + ], + "source_commit": "b487f3866cc13dd971819cb902517a6a50282404" +} diff --git a/blueprints/convergence-practice/arb-trace2code/upstream-bm25-summary.json b/blueprints/convergence-practice/arb-trace2code/upstream-bm25-summary.json new file mode 100644 index 000000000..3e15a7d72 --- /dev/null +++ b/blueprints/convergence-practice/arb-trace2code/upstream-bm25-summary.json @@ -0,0 +1,96 @@ +{ + "candidate_filter": "all_files", + "evaluated": 101, + "keep_list": null, + "metrics": { + "overall": { + "F0.5@10": 0.048909254429210854, + "F0.5@20": 0.03767778906731239, + "F0.5@5": 0.06249848587343206, + "MRR": 0.16384832699974178, + "Precision@10": 0.0405940594059406, + "Precision@20": 0.030693069306930693, + "Precision@5": 0.05346534653465347, + "Recall@10": 0.3217821782178218, + "Recall@20": 0.49339933993399343, + "Recall@5": 0.22277227722772278, + "block_f0.5@8k": 0.010993401463698494, + "block_f1@8k": 0.011134684897061134, + "block_precision@8k": 0.011526976873511525, + "block_recall@8k": 0.020702070207020702, + "block_samples": 1.0, + "context_efficiency@8k": 0.05067640769153604, + "context_pollution_tokens@8k": 4903.544554455446, + "coverage_auc@20": 0.3213696369636964, + "gold_blocks": 4.623762376237623, + "gold_coverage@8k": 0.12046204620462046, + "gold_token_ratio@8k": 0.05067640769153604, + "hard_negative_hits@10": 1.900990099009901, + "hard_negative_hits@20": 2.3564356435643563, + "hard_negative_hits@5": 1.1485148514851484, + "irrelevant_files@10": 9.594059405940595, + "irrelevant_files@20": 19.386138613861387, + "irrelevant_files@5": 4.732673267326732, + "line_f0.5@8k": 0.006755857409932814, + "line_f1@8k": 0.009432025250600162, + "line_gold_count": 50.21782178217822, + "line_overlap_count": 1.0297029702970297, + "line_precision@8k": 0.005700437738366717, + "line_predicted_count": 199.44554455445544, + "line_recall@8k": 0.05763819675672622, + "line_samples": 1.0, + "matched_blocks": 0.04950495049504951, + "predicted_blocks": 3.7524752475247523, + "redundancy@8k": 0.24071376332568756, + "samples": 101 + }, + "trace2code": { + "F0.5@10": 0.048909254429210854, + "F0.5@20": 0.03767778906731239, + "F0.5@5": 0.06249848587343206, + "MRR": 0.16384832699974178, + "Precision@10": 0.0405940594059406, + "Precision@20": 0.030693069306930693, + "Precision@5": 0.05346534653465347, + "Recall@10": 0.3217821782178218, + "Recall@20": 0.49339933993399343, + "Recall@5": 0.22277227722772278, + "block_f0.5@8k": 0.010993401463698494, + "block_f1@8k": 0.011134684897061134, + "block_precision@8k": 0.011526976873511525, + "block_recall@8k": 0.020702070207020702, + "block_samples": 1.0, + "context_efficiency@8k": 0.05067640769153604, + "context_pollution_tokens@8k": 4903.544554455446, + "coverage_auc@20": 0.3213696369636964, + "gold_blocks": 4.623762376237623, + "gold_coverage@8k": 0.12046204620462046, + "gold_token_ratio@8k": 0.05067640769153604, + "hard_negative_hits@10": 1.900990099009901, + "hard_negative_hits@20": 2.3564356435643563, + "hard_negative_hits@5": 1.1485148514851484, + "irrelevant_files@10": 9.594059405940595, + "irrelevant_files@20": 19.386138613861387, + "irrelevant_files@5": 4.732673267326732, + "line_f0.5@8k": 0.006755857409932814, + "line_f1@8k": 0.009432025250600162, + "line_gold_count": 50.21782178217822, + "line_overlap_count": 1.0297029702970297, + "line_precision@8k": 0.005700437738366717, + "line_predicted_count": 199.44554455445544, + "line_recall@8k": 0.05763819675672622, + "line_samples": 1.0, + "matched_blocks": 0.04950495049504951, + "predicted_blocks": 3.7524752475247523, + "redundancy@8k": 0.24071376332568756, + "samples": 101 + } + }, + "mode": "corpus", + "ranker": "bm25", + "runtime": { + "progress": false, + "wall_time_seconds": 9.94414224999855 + }, + "skipped": {} +} diff --git a/blueprints/convergence-practice/arb-trace2code/upstream-lexical-summary.json b/blueprints/convergence-practice/arb-trace2code/upstream-lexical-summary.json new file mode 100644 index 000000000..723dbc4f3 --- /dev/null +++ b/blueprints/convergence-practice/arb-trace2code/upstream-lexical-summary.json @@ -0,0 +1,96 @@ +{ + "candidate_filter": "all_files", + "evaluated": 101, + "keep_list": null, + "metrics": { + "overall": { + "F0.5@10": 0.06710573010141141, + "F0.5@20": 0.05050495254246749, + "F0.5@5": 0.09498465374487759, + "MRR": 0.20745277382862923, + "Precision@10": 0.05544554455445545, + "Precision@20": 0.04108910891089109, + "Precision@5": 0.0811881188118812, + "Recall@10": 0.48184818481848185, + "Recall@20": 0.6963696369636964, + "Recall@5": 0.3432343234323432, + "block_f0.5@8k": 0.0069573546122359605, + "block_f1@8k": 0.008171690408477468, + "block_precision@8k": 0.00861032531824611, + "block_recall@8k": 0.02856785678567857, + "block_samples": 1.0, + "context_efficiency@8k": 0.04563379103607761, + "context_pollution_tokens@8k": 6605.108910891089, + "coverage_auc@20": 0.46947194719471946, + "gold_blocks": 4.623762376237623, + "gold_coverage@8k": 0.12871287128712872, + "gold_token_ratio@8k": 0.04563379103607761, + "hard_negative_hits@10": 1.9702970297029703, + "hard_negative_hits@20": 2.504950495049505, + "hard_negative_hits@5": 1.4059405940594059, + "irrelevant_files@10": 9.445544554455445, + "irrelevant_files@20": 19.178217821782177, + "irrelevant_files@5": 4.594059405940594, + "line_f0.5@8k": 0.004842276083344289, + "line_f1@8k": 0.006381726345130579, + "line_gold_count": 50.21782178217822, + "line_overlap_count": 0.7326732673267327, + "line_precision@8k": 0.004329431616204497, + "line_predicted_count": 239.36633663366337, + "line_recall@8k": 0.040633869682367756, + "line_samples": 1.0, + "matched_blocks": 0.07920792079207921, + "predicted_blocks": 23.217821782178216, + "redundancy@8k": 0.7843470572627743, + "samples": 101 + }, + "trace2code": { + "F0.5@10": 0.06710573010141141, + "F0.5@20": 0.05050495254246749, + "F0.5@5": 0.09498465374487759, + "MRR": 0.20745277382862923, + "Precision@10": 0.05544554455445545, + "Precision@20": 0.04108910891089109, + "Precision@5": 0.0811881188118812, + "Recall@10": 0.48184818481848185, + "Recall@20": 0.6963696369636964, + "Recall@5": 0.3432343234323432, + "block_f0.5@8k": 0.0069573546122359605, + "block_f1@8k": 0.008171690408477468, + "block_precision@8k": 0.00861032531824611, + "block_recall@8k": 0.02856785678567857, + "block_samples": 1.0, + "context_efficiency@8k": 0.04563379103607761, + "context_pollution_tokens@8k": 6605.108910891089, + "coverage_auc@20": 0.46947194719471946, + "gold_blocks": 4.623762376237623, + "gold_coverage@8k": 0.12871287128712872, + "gold_token_ratio@8k": 0.04563379103607761, + "hard_negative_hits@10": 1.9702970297029703, + "hard_negative_hits@20": 2.504950495049505, + "hard_negative_hits@5": 1.4059405940594059, + "irrelevant_files@10": 9.445544554455445, + "irrelevant_files@20": 19.178217821782177, + "irrelevant_files@5": 4.594059405940594, + "line_f0.5@8k": 0.004842276083344289, + "line_f1@8k": 0.006381726345130579, + "line_gold_count": 50.21782178217822, + "line_overlap_count": 0.7326732673267327, + "line_precision@8k": 0.004329431616204497, + "line_predicted_count": 239.36633663366337, + "line_recall@8k": 0.040633869682367756, + "line_samples": 1.0, + "matched_blocks": 0.07920792079207921, + "predicted_blocks": 23.217821782178216, + "redundancy@8k": 0.7843470572627743, + "samples": 101 + } + }, + "mode": "corpus", + "ranker": "lexical", + "runtime": { + "progress": false, + "wall_time_seconds": 10.135793833000207 + }, + "skipped": {} +} diff --git a/blueprints/convergence-practice/arb-trace2code/upstream-validation.json b/blueprints/convergence-practice/arb-trace2code/upstream-validation.json new file mode 100644 index 000000000..d7ba51e2d --- /dev/null +++ b/blueprints/convergence-practice/arb-trace2code/upstream-validation.json @@ -0,0 +1,8 @@ +[ + { + "path": "data/benchmark/v2_trace2code/samples.jsonl", + "total": 101, + "invalid": 0, + "errors_by_id": {} + } +] diff --git a/blueprints/convergence-practice/local-fixture/corpus-manifest.json b/blueprints/convergence-practice/local-fixture/corpus-manifest.json new file mode 100644 index 000000000..2cd33c5e0 --- /dev/null +++ b/blueprints/convergence-practice/local-fixture/corpus-manifest.json @@ -0,0 +1,110 @@ +{ + "schema_version": 1, + "corpus_id": "public-convergence-docs-2026-09-20-v1", + "created_at": "2026-09-20T02:06:03.632262+00:00", + "source_repository_label": "native-agent-stack public reference", + "base_commit": "bf99d340f798bf72cc3104619edc482882eb423d", + "selection": "Fifteen directly read public documentation files selected before evaluation, with source bytes fixed by the base commit. No retrieval or ranking selected these files.", + "document_count": 15, + "path_order": "case-sensitive canonical repository-relative paths in ascending order", + "files": [ + { + "path": "adoption/README.md", + "source_sha256": "5e554f816905070b474aaa8e177fbc7b4ed81e017019fb3067ec08ca67b6c568", + "utf8_bytes": 9783, + "media_type": "text/markdown" + }, + { + "path": "adoption/update.md", + "source_sha256": "4aedc083e4665bec7131263e36b81390e4011c660b3c78df183ddc988021aadd", + "utf8_bytes": 8234, + "media_type": "text/markdown" + }, + { + "path": "blueprints/us-equities/acceptance-wave/research-protocol.md", + "source_sha256": "a4bbb4b1682e84c7324b903f90ececce3a3f7c547b445517089607794ded4f92", + "utf8_bytes": 6449, + "media_type": "text/markdown" + }, + { + "path": "blueprints/us-equities/authenticated-data/README.md", + "source_sha256": "76d2439e456aacca301013d1993bde75e42fd6ffa0f77990eb2e60686a961666", + "utf8_bytes": 8763, + "media_type": "text/markdown" + }, + { + "path": "blueprints/us-equities/data-readiness/README.md", + "source_sha256": "77ac8fa53c3fca2ede3824a80a17d448a37f28d36e872c4d619af2c4be55642a", + "utf8_bytes": 5426, + "media_type": "text/markdown" + }, + { + "path": "blueprints/us-equities/memory-lifecycle/README.md", + "source_sha256": "309fc735617e7b939a91866301614707a713bd9e38d1b4f241722ecd5925a6cf", + "utf8_bytes": 4321, + "media_type": "text/markdown" + }, + { + "path": "blueprints/us-equities/research-evaluation/README.md", + "source_sha256": "5f6e4b5cc7338e9f5fc3edeb0d2d5092f1c2ba779d0243ed3ba29da76a80edde", + "utf8_bytes": 8589, + "media_type": "text/markdown" + }, + { + "path": "blueprints/us-equities/research-runtime/README.md", + "source_sha256": "26941ae2c64ec137c3a9dc97e8d766d93cf6bdcb103ee290f871fd0a181f2757", + "utf8_bytes": 9587, + "media_type": "text/markdown" + }, + { + "path": "blueprints/us-equities/state-recovery/README.md", + "source_sha256": "bc48a54ba5bc9a0f7c8128cfba5386531b26c12bbe51cc97002ce2148d50e8c7", + "utf8_bytes": 1381, + "media_type": "text/markdown" + }, + { + "path": "blueprints/us-equities/state-recovery/memory/README.md", + "source_sha256": "c58fd2a181feb488281c5a0a57fe7b6227f085ad44c41c53f539632d26a34965", + "utf8_bytes": 8477, + "media_type": "text/markdown" + }, + { + "path": "blueprints/us-equities/state-recovery/qdrant/README.md", + "source_sha256": "13f5ce8292fcec69e3daa85859c800aecb4e529bf1f94391dc883129623febe0", + "utf8_bytes": 5252, + "media_type": "text/markdown" + }, + { + "path": "blueprints/us-equities/worker-supervision/README.md", + "source_sha256": "a0b55860e727e2bd7e631f7070e50bf5aa6c98974fd94dea7ac72362d1aca695", + "utf8_bytes": 5670, + "media_type": "text/markdown" + }, + { + "path": "docs/evidence.md", + "source_sha256": "a5d81d728d58571ac7c88e6aafb26fcb8707717541a92e55eb39d58d8ad55353", + "utf8_bytes": 9129, + "media_type": "text/markdown" + }, + { + "path": "observability/desktop-restart.md", + "source_sha256": "39e4ee70ea568d8f30182715566c287abc0b6b06c6176634c93fdae2e7dbc711", + "utf8_bytes": 6238, + "media_type": "text/markdown" + }, + { + "path": "observability/session-e2e.md", + "source_sha256": "843e8f4dff03e7f81d260b060ed3f140c2f6574555680beb84f3f7c0b2888712", + "utf8_bytes": 7704, + "media_type": "text/markdown" + } + ], + "hashfreeze": { + "algorithm": "sha256", + "canonicalization": "UTF-8 JSON, sorted keys, compact separators, ensure_ascii=false", + "payload_scope": "entire top-level object excluding hashfreeze", + "payload_sha256": "f104af0112d6c6307a5bc7e08e3f9e65835e9efef35b4e69c098f948058bae83", + "frozen_before_first_search_or_ranking_run": true, + "change_policy": "Never edit scored questions or labels in place; create a new version and retain prior results." + } +} diff --git a/blueprints/convergence-practice/local-fixture/evaluation-receipt.json b/blueprints/convergence-practice/local-fixture/evaluation-receipt.json new file mode 100644 index 000000000..bca64ff22 --- /dev/null +++ b/blueprints/convergence-practice/local-fixture/evaluation-receipt.json @@ -0,0 +1,127 @@ +{ + "artifact_hashes": { + "evaluation.md": "695e61ebe524d0559280a7f3c9247c6e613355215387678ce76fd88411d76b9a", + "rankings.json": "35fc05c3483fa5f4a674ee190c543748c1f72d589d9b8f0ed83c2c667f8cd36e" + }, + "checks": { + "all_15_git_source_blobs_hash_and_byte_length_match": true, + "changed_fixture_and_corpus_bytes_rejected": true, + "check_kind": "Ad hoc invariant and arithmetic checks for this bounded documented recipe; no new general-purpose adapter test suite.", + "complete_normalized_file_bodies_preserved": true, + "each_ranking_has_all_15_unique_source_paths": true, + "fixture_and_manifest_exact_hashes": true, + "independent_recomputation_from_ranked_paths_matches_all_cases_and_macro_means": true, + "known_metric_examples_passed": true, + "unsafe_paths_rejected": true, + "upstream_corpus_baseline_tests": "15 passed at exact release; see ../arb-trace2code/receipt.json#/checks/upstream_tests" + }, + "configuration": { + "bm25": { + "b": 0.75, + "k1": 1.5 + }, + "candidate_documents_per_query": 15, + "chunk_parameter_selection": "Maximum raw frozen document character length, approved before scoring; same for both rankers.", + "chunks": 15, + "cutoffs_predeclared": [ + 1, + 3, + 5 + ], + "full_body_normalization": "Native CRLF/CR to LF and surrounding whitespace strip only.", + "max_chunk_chars": 9777, + "models_called": 0, + "mrr": "Full ranking, not truncated at any k.", + "native_apis": [ + "corpus.chunks_for_file", + "baseline.rank_chunks_for_ranker", + "baseline.unique_ranked_paths", + "baseline.recall_at", + "baseline.sample_metrics" + ], + "native_indexes_changed": 0, + "negative_policy": "Four no-gold cases ranked and retained, excluded from positive metrics; no threshold, abstention policy or accuracy.", + "positive_aggregation": "Macro mean over 20 one-gold-file positive queries.", + "provider_calls": 0, + "query_transform": "None", + "ranker_order": "sequential lexical then bm25", + "rankers": [ + "lexical", + "bm25" + ] + }, + "inputs": { + "base_commit": "bf99d340f798bf72cc3104619edc482882eb423d", + "corpus_documents": 15, + "corpus_hashes": "corpus-manifest.json#/files", + "corpus_manifest_sha256": "7bb986160b332e205208b6767b95b6f019423ffed1f50843534ff7e4272fdded", + "expected_abstention_queries": 4, + "fixture_sha256": "aea5e9616fb19eb1d2fd58ecfadf88ff087d7d2e79b4a70bc9f88207af9d0402", + "frozen_before_first_rank": true, + "positive_queries": 20, + "queries": 24, + "queries_verbatim": true, + "source_authored_not_external_holdout": true, + "source_repository": "https://github.com/seathatflowsinourveins/native-agent-stack" + }, + "kind": "executed_frozen_public_document_retrieval_fixture", + "limits": [ + "Source-authored 24-question integration fixture over only 15 selected public docs; not an external held-out benchmark.", + "Four negative cases have rankings only, no abstention or answer-quality metric.", + "Single warm process ranker-loop observations exclude source loading and are not a latency benchmark.", + "No claim about rg, QMD, SocratiCode, a native memory service or production tool acceptance.", + "No inference, tokenizer cost, provider usage, scientific significance or default-promotion claim." + ], + "observed_date_utc": "2026-09-20", + "replay": { + "command_template": "PYTHONPATH=upstream/src FIXTURE_DIRECTORY=${FIXTURE_DIRECTORY} STACK_REPOSITORY=${STACK_REPOSITORY} REPLAY_OUTPUT=${REPLAY_OUTPUT} runtime/bin/python local_doc_replay.py", + "document": "evaluation.md", + "environment": "Isolated stdlib-only virtual environment, no installed packages.", + "execution_exit_code": 0, + "python": "3.14.7", + "recipe_source_sha256": "557b8cf4b3123d3e0c3dfc9d688950d75cc446f3bc162825eff24708fa78ab4d" + }, + "results": { + "bm25": { + "abstention_accuracy": null, + "abstention_policy": null, + "macro_metrics": { + "MRR": 0.9666666666666666, + "Recall@1": 0.95, + "Recall@3": 1.0, + "Recall@5": 1.0 + }, + "negative_rankings_returned": 4, + "negative_samples_excluded_from_positive_metrics": 4, + "positive_samples": 20, + "wall_seconds": 0.07384333299705759 + }, + "lexical": { + "abstention_accuracy": null, + "abstention_policy": null, + "macro_metrics": { + "MRR": 0.9, + "Recall@1": 0.8, + "Recall@3": 1.0, + "Recall@5": 1.0 + }, + "negative_rankings_returned": 4, + "negative_samples_excluded_from_positive_metrics": 4, + "positive_samples": 20, + "wall_seconds": 0.06728295799985062 + } + }, + "schema_version": 1, + "source": { + "commit": "b487f3866cc13dd971819cb902517a6a50282404", + "files": { + "src/agent_retrieval_bench/baseline.py": "5a82f9a087eb90f9ada2ad6b6047fe8b9781b013749f3d176b1cbdbe7b7f05d1", + "src/agent_retrieval_bench/corpus.py": "6ddde17c0654a9f9086f6054c1ffaadbed00fc3a0a4af8c71dbce57820c69f52" + }, + "git_source_tree": "0fd0466fb3e24c40c0bdf9468872bea9c9a61fa1", + "release": "v0.2.1", + "repository": "https://github.com/eyuansu62/agent-retrieval-bench", + "runtime_dependencies": [], + "unmodified": true + } +} diff --git a/blueprints/convergence-practice/local-fixture/evaluation.md b/blueprints/convergence-practice/local-fixture/evaluation.md new file mode 100644 index 000000000..f8c1c5278 --- /dev/null +++ b/blueprints/convergence-practice/local-fixture/evaluation.md @@ -0,0 +1,215 @@ +# Frozen public-document retrieval replay + +This is an executed integration fixture using the unchanged Agent Retrieval Bench +v0.2.1 lexical and BM25 APIs, not a held-out external benchmark or a live native +search acceptance check. The independently authored fixture selected 15 public +repository documents and froze 24 natural questions before ranking. The author +could read the documents; neither queries, labels nor cutoffs were tuned after +seeing these results. All source bytes come from the immutable native-agent-stack +commit `bf99d340f798bf72cc3104619edc482882eb423d`. + +| Native ranker | Positive denominator | Recall@1 | Recall@3 | Recall@5 | Full-ranking MRR | +| --- | ---: | ---: | ---: | ---: | ---: | +| lexical | 20 | 0.80 | 1.00 | 1.00 | 0.90 | +| BM25 | 20 | 0.95 | 1.00 | 1.00 | 0.9666666667 | + +Recall is the macro mean of each case's fraction of exact gold file paths found +within the first k unique file paths. Each of these 20 positives has one gold +file. MRR uses the first gold file in the full ranking. Cutoffs 1, 3 and 5 were +chosen before scoring this small corpus; Recall@20 would be trivial for 15 files. +There is no significance or general superiority claim. The paired mean differences +(BM25 minus lexical) are +0.15 Recall@1 and +0.0666666667 MRR, with zero difference +at Recall@3 and Recall@5. + +Four expected-abstention cases have no gold file and are explicitly excluded from +the positive metric denominator. Both native rankers returned all 15 files for +each of those four cases. Their complete rankings are retained, with `metrics: +null`. There is no abstention policy, score threshold, abstention accuracy, answer +generation or answer-quality assessment. The fixture's `boundary_files` are +explanatory labels and never become gold targets. + +The maximum raw document character length, 9,777, was selected as the chunk size +before scoring. Native `chunks_for_file` produced exactly 15 file chunks and the +recipe checked that every normalized complete file body was preserved. Queries +were passed verbatim to the native rankers. BM25 retains native k1=1.5 and b=0.75; +lexical retains its native path/basename/symbol bonuses and unique-token +normalization. Algorithm details and exact release provenance are in the +[external evaluation](../arb-trace2code/README.md). + +The two sequential ranker-loop observations were approximately 0.0673 seconds +(lexical) and 0.0738 seconds (BM25), covering their 24 rankings and metric work. +These are single warm process observations, excluding source loading, with no +controlled latency comparison. No model, provider, active index, daemon, native +account, private document or durable memory was accessed or changed. + +`rankings.json` retains 48 full file rankings and the exact per-case +and macro metrics. `evaluation-receipt.json` records source/input/code hashes, +parameters and checks. The frozen questions and corpus manifest are alongside +this document; no source document bodies are copied into this packet. The very +small, source-authored corpus can establish executable integration and inspectable +retrieval behavior, not real-user retrieval quality, abstention, production search +acceptance, token savings or a new installed default. + +## Reproduce with the native APIs + +Use the pinned `upstream/` checkout and dependency-free `runtime/` environment from +the [external replay recipe](../arb-trace2code/README.md), in the same scratch +working directory. Confirm the checkout is clean and its HEAD is +`b487f3866cc13dd971819cb902517a6a50282404`. No optional dependencies are needed. +Point `STACK_REPOSITORY` to a clone of +[the public reference repository](https://github.com/seathatflowsinourveins/native-agent-stack) +that contains the frozen commit and this published fixture. Set +`FIXTURE_DIRECTORY` to its `blueprints/convergence-practice/local-fixture` directory +and `REPLAY_OUTPUT` to a new output file in scratch. The source working tree's +current branch is immaterial: each document is read with `git show` at the fixed +commit and checked against its exact frozen hash. + +Save the following fixture-specific recipe verbatim as `local_doc_replay.py` in +scratch. Its SHA256 is +`557b8cf4b3123d3e0c3dfc9d688950d75cc446f3bc162825eff24708fa78ab4d`. +It calls the upstream chunker, rankers and metrics directly; it is not a reusable +adapter or a replacement retrieval implementation. Subprocess input is an +argument array, source paths are canonical relative paths, and a hash or corpus +mismatch stops the run. The output uses exclusive creation to preserve earlier +observations. + +```python +"""Replay one frozen documentation fixture with unmodified ARB v0.2.1 APIs.""" + +import hashlib +import json +import os +import subprocess +import time +from pathlib import Path, PurePosixPath + +from agent_retrieval_bench.baseline import ( + rank_chunks_for_ranker, + recall_at, + sample_metrics, + unique_ranked_paths, +) +from agent_retrieval_bench.corpus import chunks_for_file + +BASE = "bf99d340f798bf72cc3104619edc482882eb423d" +FIXTURE_SHA = "aea5e9616fb19eb1d2fd58ecfadf88ff087d7d2e79b4a70bc9f88207af9d0402" +MANIFEST_SHA = "7bb986160b332e205208b6767b95b6f019423ffed1f50843534ff7e4272fdded" +REPO = "seathatflowsinourveins/native-agent-stack" + + +def checked_bytes(raw, expected): + if hashlib.sha256(raw).hexdigest() != expected: + raise ValueError("Frozen input hash mismatch") + return raw + + +def source_path(value): + path = PurePosixPath(value) + if (path.is_absolute() or ".." in path.parts or ":" in value + or "\\" in value or value != path.as_posix()): + raise ValueError("Source path is not a canonical relative path") + return value + + +def main(): + fixture_dir = Path(os.environ["FIXTURE_DIRECTORY"]) + source_repo = Path(os.environ["STACK_REPOSITORY"]) + output = Path(os.environ["REPLAY_OUTPUT"]) + fixture = json.loads(checked_bytes((fixture_dir / "fixture.json").read_bytes(), FIXTURE_SHA)) + manifest = json.loads(checked_bytes( + (fixture_dir / "corpus-manifest.json").read_bytes(), MANIFEST_SHA)) + if fixture["base_commit"] != BASE or manifest["base_commit"] != BASE: + raise ValueError("Frozen base commit mismatch") + documents = [] + for item in manifest["files"]: + path = source_path(item["path"]) + raw = subprocess.run( + ["git", "-C", str(source_repo), "show", f"{BASE}:{path}"], + check=True, stdout=subprocess.PIPE, stderr=subprocess.PIPE, timeout=10, + ).stdout + checked_bytes(raw, item["source_sha256"]) + if len(raw) != item["utf8_bytes"]: + raise ValueError("Frozen document length mismatch") + documents.append((path, raw.decode("utf-8"))) + paths = {path for path, _ in documents} + if len(documents) != 15 or len(paths) != 15 or len(fixture["queries"]) != 24: + raise ValueError("Frozen corpus or query count mismatch") + maximum = max(len(text) for _, text in documents) + chunks = [] + for path, text in documents: + native = chunks_for_file(REPO, BASE, path, text, max_chunk_chars=maximum) + if native[0]["text"] != text.replace("\r\n", "\n").replace("\r", "\n").strip(): + raise ValueError("Native chunker did not preserve full normalized file body") + chunks.extend(native) + rows = [] + timings = {} + for ranker in ("lexical", "bm25"): + start = time.monotonic() + for query in fixture["queries"]: + gold = set(query["expected_files"]) + if not gold.issubset(paths) or query["expected_abstention"] != (not gold): + raise ValueError("Frozen gold labels are inconsistent") + ranked = rank_chunks_for_ranker(query["query"], chunks, ranker) + files = unique_ranked_paths(ranked) + if set(files) != paths or len(files) != 15: + raise ValueError("Native ranking changed the candidate universe") + metrics = None + if gold: + native_metrics = sample_metrics(sorted(gold), ranked) + metrics = {f"Recall@{k}": recall_at(gold, files, k) for k in (1, 3, 5)} + metrics["MRR"] = native_metrics["MRR"] + rows.append({ + "sample_id": query["id"], "ranker": ranker, + "expected_abstention": query["expected_abstention"], + "gold_files": sorted(gold), "ranked_files": files, + "metrics": metrics, + }) + timings[ranker] = time.monotonic() - start + summaries = {} + for ranker in ("lexical", "bm25"): + positives = [row for row in rows if row["ranker"] == ranker and row["metrics"]] + negatives = [row for row in rows if row["ranker"] == ranker + and row["expected_abstention"]] + if len(positives) != 20 or len(negatives) != 4: + raise ValueError("Frozen positive or negative denominator mismatch") + summaries[ranker] = { + "positive_samples": 20, "negative_samples_excluded_from_positive_metrics": 4, + "macro_metrics": {key: sum(row["metrics"][key] for row in positives) / 20 + for key in ("Recall@1", "Recall@3", "Recall@5", "MRR")}, + "negative_rankings_returned": sum(bool(row["ranked_files"]) for row in negatives), + "abstention_policy": None, "abstention_accuracy": None, + "wall_seconds": timings[ranker], + } + result = { + "schema_version": 1, "base_commit": BASE, "fixture_sha256": FIXTURE_SHA, + "corpus_manifest_sha256": MANIFEST_SHA, + "documents": 15, "chunks": len(chunks), "max_chunk_chars": maximum, + "full_body_normalization": "Native CRLF/CR to LF and surrounding whitespace strip only.", + "query_transform": "None: original fixture query string passed verbatim to native ranker.", + "summaries": summaries, "results": rows, + } + output.parent.mkdir(parents=True, exist_ok=True) + with output.open("x") as handle: + handle.write(json.dumps(result, indent=2, sort_keys=True) + "\n") + print(json.dumps({key: result[key] for key in ("documents", "chunks", "max_chunk_chars", + "summaries")}, indent=2)) + + +if __name__ == "__main__": + main() +``` + +After setting the three environment variables above, run: + +```sh +PYTHONPATH=upstream/src runtime/bin/python local_doc_replay.py +``` + +The executed command used those same three environment variables with private +scratch locations. Exact locations are deliberately omitted from public evidence. +Compare rankings and metrics, allowing the recorded timing values to differ. +The receipt retains the exact source file hash of this executed recipe. The input +guards were checked against changed fixture bytes, changed corpus bytes and +traversal paths before execution; known Recall and MRR examples and independent +rank-based recomputation also passed. This is a bounded documented analysis +recipe, not a new installed component or a general-purpose adapter test suite. diff --git a/blueprints/convergence-practice/local-fixture/fixture.json b/blueprints/convergence-practice/local-fixture/fixture.json new file mode 100644 index 000000000..2f70582b1 --- /dev/null +++ b/blueprints/convergence-practice/local-fixture/fixture.json @@ -0,0 +1,649 @@ +{ + "schema_version": 1, + "fixture_id": "public-convergence-questions-2026-09-20-v1", + "created_at": "2026-09-20T02:06:03.632262+00:00", + "base_commit": "bf99d340f798bf72cc3104619edc482882eb423d", + "corpus_manifest": "corpus-manifest.json", + "corpus_manifest_sha256": "7bb986160b332e205208b6767b95b6f019423ffed1f50843534ff7e4272fdded", + "question_count": 24, + "answerable_count": 20, + "expected_abstention_count": 4, + "authorship": { + "method": "One agent authored natural user questions, relevance judgments and exact support quotations by directly reading public source documents. No model API calls or search/ranking runs were performed during authoring.", + "queries_authored_before_first_evaluation": true, + "search_or_ranking_executed_by_author": false, + "independent_of_evaluation_outputs": true, + "external_real_user_benchmark": false, + "blinded_external_relevance_judgments": false, + "prior_fixture_reviewed": "blueprints/us-equities/retrieval-evaluation/fixture.json", + "prior_fixture_question_count": 12, + "prior_fixture_status": "Previously scored historical evidence; not reused or relabeled as a fresh held-out benchmark." + }, + "evaluation_contract": { + "query_text": "Use query verbatim; it is the full natural question, without post-score lexical rewrites.", + "document_id": "Exact case-sensitive repository-relative path; no basename or suffix matching.", + "answerable_retrieval": "Score expected_files only. Report recall at fixed cutoffs and rank separately from answer correctness.", + "abstention_retrieval": "expected_files is intentionally empty for abstentions; boundary_files are explanatory sources, not positive answer targets. Do not count a retrieved boundary as an answerable-document hit.", + "answer_judgment": "Assess support, source citation and adherence to reference_answer boundaries separately; do not infer answer quality from source retrieval alone.", + "corpus_scope": "Index only the exact fifteen source files in corpus-manifest.json. Do not index this question file, manifests, reports, old evaluation fixtures or other repository files.", + "failure_policy": "Keep backend errors distinct from empty results and misses. Retain every original question/result; document any new fixture version separately." + }, + "limitations": [ + "Small source-authored repository integration fixture, not an independent real-user external benchmark or general retrieval-quality estimate.", + "Questions and labels were derived with source documents visible; topic overlap with earlier repository guidance is unavoidable and does not create a fresh external holdout.", + "One author selected relevance labels; positive gold paths are reviewed intended sources, not exhaustive judgments of every relevant passage.", + "No search/ranking performance, answer accuracy, embedding quality, latency, model cost, token savings, platform portability or live runtime acceptance has been measured by authoring this fixture.", + "The abstention labels apply to the requested specifics in this frozen public corpus, not universal unanswerability.", + "Source updates must not silently alter the corpus; preserve this base commit and hashes for comparisons." + ], + "queries": [ + { + "id": "credentials-01", + "category": "native_credentials_host_evidence", + "query": "After cloning the stack onto another machine, how should native sign-in and tool activation be checked across Linux clients and Desktop?", + "question": "After cloning the stack onto another machine, how should native sign-in and tool activation be checked across Linux clients and Desktop?", + "expected_behavior": "answer_with_sources", + "expected_abstention": false, + "expected_files": [ + "adoption/README.md" + ], + "boundary_files": [], + "relevance_rationale": "The ordered adoption instructions state both the required native account action and the separate client activation scopes.", + "reference_answer": "Sign in natively on the destination, then verify discovery and a useful scoped call for each relevant client separately; do not copy account stores.", + "source_evidence": [ + { + "path": "adoption/README.md", + "source_sha256": "5e554f816905070b474aaa8e177fbc7b4ed81e017019fb3067ec08ca67b6c568", + "utf8_bytes": 9783, + "line_start": 33, + "line_end": 33, + "snippet": "Use native sign-in on the target host. Then verify the actual client's tool discovery and one bounded useful call. Treat Linux Codex, Linux Claude and Desktop as separate client scopes. No authentication store is copied.", + "support_role": "answer_support" + } + ] + }, + { + "id": "credentials-02", + "category": "native_credentials_host_evidence", + "query": "The prerequisite report returned zero. Does that tell me my provider account and running services are ready?", + "question": "The prerequisite report returned zero. Does that tell me my provider account and running services are ready?", + "expected_behavior": "answer_with_sources", + "expected_abstention": false, + "expected_files": [ + "adoption/README.md" + ], + "boundary_files": [], + "relevance_rationale": "This paragraph defines the exact exit-code contract and the probes it excludes.", + "reference_answer": "No. Zero confirms selected prerequisites only; account and live-service readiness need separate checks.", + "source_evidence": [ + { + "path": "adoption/README.md", + "source_sha256": "5e554f816905070b474aaa8e177fbc7b4ed81e017019fb3067ec08ca67b6c568", + "utf8_bytes": 9783, + "line_start": 46, + "line_end": 46, + "snippet": "A prerequisite report exits 0 when the selected prerequisites are present, 2 when they are missing, unsupported or malformed. It does not probe accounts or services.", + "support_role": "answer_support" + } + ] + }, + { + "id": "evidence-03", + "category": "native_credentials_host_evidence", + "query": "What do the published receipt checksums establish about an earlier model run, and what do they leave unverified?", + "question": "What do the published receipt checksums establish about an earlier model run, and what do they leave unverified?", + "expected_behavior": "answer_with_sources", + "expected_abstention": false, + "expected_files": [ + "docs/evidence.md" + ], + "boundary_files": [], + "relevance_rationale": "The evidence guide explicitly distinguishes artifact integrity from independent provider attestation or replay.", + "reference_answer": "They verify the published artifact bytes. They do not independently attest provider execution or constitute a newly executed model task.", + "source_evidence": [ + { + "path": "docs/evidence.md", + "source_sha256": "a5d81d728d58571ac7c88e6aafb26fcb8707717541a92e55eb39d58d8ad55353", + "utf8_bytes": 9129, + "line_start": 3, + "line_end": 3, + "snippet": "These hashes establish integrity of the published bytes; they do not independently authenticate a provider or turn a recorded result into a new run.", + "support_role": "answer_support" + } + ] + }, + { + "id": "memory-04", + "category": "memory", + "query": "If a pinned memory page has expired, can an exact-path read still return it before maintenance runs?", + "question": "If a pinned memory page has expired, can an exact-path read still return it before maintenance runs?", + "expected_behavior": "answer_with_sources", + "expected_abstention": false, + "expected_files": [ + "blueprints/us-equities/memory-lifecycle/README.md" + ], + "boundary_files": [], + "relevance_rationale": "The native lifecycle result directly covers expired-page search and exact-path access before a sweep.", + "reference_answer": "Yes in the recorded fixture. Expiry hid it from ordinary search but exact-path and include_expired reads still returned it before the explicit sweep.", + "source_evidence": [ + { + "path": "blueprints/us-equities/memory-lifecycle/README.md", + "source_sha256": "309fc735617e7b939a91866301614707a713bd9e38d1b4f241722ecd5925a6cf", + "utf8_bytes": 4321, + "line_start": 7, + "line_end": 7, + "snippet": "**Expiry is not immediate access revocation.** An already expired pinned page was hidden from ordinary search, but `include_expired` and an exact-path read returned it before the sweep.", + "support_role": "answer_support" + } + ] + }, + { + "id": "memory-05", + "category": "memory", + "query": "How did the disposable memory exercise distinguish two projects that used the same page path?", + "question": "How did the disposable memory exercise distinguish two projects that used the same page path?", + "expected_behavior": "answer_with_sources", + "expected_abstention": false, + "expected_files": [ + "blueprints/us-equities/memory-lifecycle/README.md" + ], + "boundary_files": [], + "relevance_rationale": "The reported project-routing check uses identical page paths and cross-project unique terms.", + "reference_answer": "Each scope returned its own page body, and searching for the other project's unique term returned no hits; this tests routing rather than tenant authorization.", + "source_evidence": [ + { + "path": "blueprints/us-equities/memory-lifecycle/README.md", + "source_sha256": "309fc735617e7b939a91866301614707a713bd9e38d1b4f241722ecd5925a6cf", + "utf8_bytes": 4321, + "line_start": 5, + "line_end": 5, + "snippet": "The same page path returned each project's own body; searches for the other project's unique term returned zero hits.", + "support_role": "answer_support" + } + ] + }, + { + "id": "memory-06", + "category": "memory", + "query": "Can the memory lifecycle acceptance be used as proof of market valid-time retrieval or secure erasure from backups?", + "question": "Can the memory lifecycle acceptance be used as proof of market valid-time retrieval or secure erasure from backups?", + "expected_behavior": "answer_with_sources", + "expected_abstention": false, + "expected_files": [ + "blueprints/us-equities/memory-lifecycle/README.md" + ], + "boundary_files": [], + "relevance_rationale": "This explicit boundary separates memory ingestion-time behavior and query deletion from stronger temporal and erasure claims.", + "reference_answer": "No. It demonstrates ingestion-time retrieval and removal from query results, not market valid-time semantics or secure erasure of checkpoints and backups.", + "source_evidence": [ + { + "path": "blueprints/us-equities/memory-lifecycle/README.md", + "source_sha256": "309fc735617e7b939a91866301614707a713bd9e38d1b4f241722ecd5925a6cf", + "utf8_bytes": 4321, + "line_start": 28, + "line_end": 28, + "snippet": "The acceptance covers project routing, not tenant authorization; ingestion-time retrieval, not market valid-time data or historical ranking; search deletion, not secure erasure of Git checkpoints/backups.", + "support_role": "answer_support" + } + ] + }, + { + "id": "context-07", + "category": "context_usage", + "query": "If the Context Mode doctor lists registered hooks, does the restart acceptance prove every lifecycle hook actually fired?", + "question": "If the Context Mode doctor lists registered hooks, does the restart acceptance prove every lifecycle hook actually fired?", + "expected_behavior": "answer_with_sources", + "expected_abstention": false, + "expected_files": [ + "observability/desktop-restart.md" + ], + "boundary_files": [], + "relevance_rationale": "The restart acceptance retains this exact distinction between registration and executed lifecycle evidence.", + "reference_answer": "No. Registration alone does not prove that every hook executed; relevant lifecycle behavior requires its own evidence.", + "source_evidence": [ + { + "path": "observability/desktop-restart.md", + "source_sha256": "39e4ee70ea568d8f30182715566c287abc0b6b06c6176634c93fdae2e7dbc711", + "utf8_bytes": 6238, + "line_start": 105, + "line_end": 105, + "snippet": "Registered hooks are not proof that every lifecycle hook was exercised.", + "support_role": "answer_support" + } + ] + }, + { + "id": "context-08", + "category": "context_usage", + "query": "How should an empty SDK receipt panel be interpreted when there are no records in the selected time range?", + "question": "How should an empty SDK receipt panel be interpreted when there are no records in the selected time range?", + "expected_behavior": "answer_with_sources", + "expected_abstention": false, + "expected_files": [ + "observability/session-e2e.md" + ], + "boundary_files": [], + "relevance_rationale": "The observation guide states the missing-data semantics and its blind spot for non-emitting tasks.", + "reference_answer": "Treat it as no matching data, not zero usage. The panel also cannot account for tasks that never emitted a receipt.", + "source_evidence": [ + { + "path": "observability/session-e2e.md", + "source_sha256": "843e8f4dff03e7f81d260b060ed3f140c2f6574555680beb84f3f7c0b2888712", + "utf8_bytes": 7704, + "line_start": 106, + "line_end": 107, + "snippet": "Empty data remains “No matching data,” not zero usage. It cannot identify tasks\nthat never emitted a receipt.", + "support_role": "answer_support" + } + ] + }, + { + "id": "context-09", + "category": "context_usage", + "query": "What local condition caused Context Mode to work in an interactive shell but fail to start from Desktop in the recorded restart investigation?", + "question": "What local condition caused Context Mode to work in an interactive shell but fail to start from Desktop in the recorded restart investigation?", + "expected_behavior": "answer_with_sources", + "expected_abstention": false, + "expected_files": [ + "observability/desktop-restart.md" + ], + "boundary_files": [], + "relevance_rationale": "The root-cause section records a process-environment executable-resolution mismatch rather than missing credentials or a plugin rewrite.", + "reference_answer": "Desktop could not resolve the installed Linux node executable on its process path, although interactive shells could. This is a historical host-specific diagnosis.", + "source_evidence": [ + { + "path": "observability/desktop-restart.md", + "source_sha256": "39e4ee70ea568d8f30182715566c287abc0b6b06c6176634c93fdae2e7dbc711", + "utf8_bytes": 6238, + "line_start": 29, + "line_end": 31, + "snippet": "Desktop's process environment could not resolve\nLinux `node`, and its native startup log reported `No such file or directory`.\nThe already installed Node 24.21.0 worked in interactive shells.", + "support_role": "answer_support" + } + ] + }, + { + "id": "worker-10", + "category": "worker_recovery", + "query": "After a local supervisor terminates a research process, what still needs checking about its remote provider request and usage?", + "question": "After a local supervisor terminates a research process, what still needs checking about its remote provider request and usage?", + "expected_behavior": "answer_with_sources", + "expected_abstention": false, + "expected_files": [ + "blueprints/us-equities/worker-supervision/README.md" + ], + "boundary_files": [], + "relevance_rationale": "The supervision adoption gate names the unresolved remote-request and usage boundary after local termination.", + "reference_answer": "Preserve partial receipts, propagate cancellation through the native client, and reconcile provider usage. A local kill alone does not establish that remote work or billing stopped.", + "source_evidence": [ + { + "path": "blueprints/us-equities/worker-supervision/README.md", + "source_sha256": "a0b55860e727e2bd7e631f7070e50bf5aa6c98974fd94dea7ac72362d1aca695", + "utf8_bytes": 5670, + "line_start": 114, + "line_end": 117, + "snippet": "Before wrapping a real research worker, select limits from actual workload\nmeasurements, propagate cancellation to the native client, preserve partial\nreceipts, and reconcile provider usage after interruption. A local kill cannot\nguarantee that a remote request stopped or that billing stopped.", + "support_role": "answer_support" + } + ] + }, + { + "id": "worker-11", + "category": "worker_recovery", + "query": "What did the successful fresh transient service establish after the timeout exercise, and did it restore unfinished application state?", + "question": "What did the successful fresh transient service establish after the timeout exercise, and did it restore unfinished application state?", + "expected_behavior": "answer_with_sources", + "expected_abstention": false, + "expected_files": [ + "blueprints/us-equities/worker-supervision/README.md" + ], + "boundary_files": [], + "relevance_rationale": "The direct-result interpretation distinguishes a fresh launch from application recovery.", + "reference_answer": "It established that a new unit could start successfully. It did not resume the interrupted application state.", + "source_evidence": [ + { + "path": "blueprints/us-equities/worker-supervision/README.md", + "source_sha256": "a0b55860e727e2bd7e631f7070e50bf5aa6c98974fd94dea7ac72362d1aca695", + "utf8_bytes": 5670, + "line_start": 99, + "line_end": 100, + "snippet": "The fresh run proves that another unit\ncan start successfully; unfinished application state was not resumed.", + "support_role": "answer_support" + } + ] + }, + { + "id": "worker-12", + "category": "worker_recovery", + "query": "Do the recorded cgroup limits demonstrate that the worker was tested under memory exhaustion, CPU saturation and a process explosion?", + "question": "Do the recorded cgroup limits demonstrate that the worker was tested under memory exhaustion, CPU saturation and a process explosion?", + "expected_behavior": "answer_with_sources", + "expected_abstention": false, + "expected_files": [ + "blueprints/us-equities/worker-supervision/README.md" + ], + "boundary_files": [], + "relevance_rationale": "The supervision receipt interpretation directly limits what the observed resource settings establish.", + "reference_answer": "No. The values establish configured limits; saturation, excessive process creation and memory-exhaustion behavior were not demonstrated.", + "source_evidence": [ + { + "path": "blueprints/us-equities/worker-supervision/README.md", + "source_sha256": "a0b55860e727e2bd7e631f7070e50bf5aa6c98974fd94dea7ac72362d1aca695", + "utf8_bytes": 5670, + "line_start": 97, + "line_end": 99, + "snippet": "The observed resource\nvalues prove limit configuration, not behavior under CPU saturation, excessive\nprocess creation or memory exhaustion.", + "support_role": "answer_support" + } + ] + }, + { + "id": "execution-13", + "category": "deterministic_execution", + "query": "Can a missing paired Astra report silently turn the Claude step into an independent review, or must that be selected explicitly?", + "question": "Can a missing paired Astra report silently turn the Claude step into an independent review, or must that be selected explicitly?", + "expected_behavior": "answer_with_sources", + "expected_abstention": false, + "expected_files": [ + "blueprints/us-equities/research-runtime/README.md" + ], + "boundary_files": [], + "relevance_rationale": "The native workflow binds paired execution to the previous result and reserves independent mode for an explicit invocation.", + "reference_answer": "Independent mode must be selected explicitly. The paired workflow requires the preceding receipt and report to match the packet and report hashes.", + "source_evidence": [ + { + "path": "blueprints/us-equities/research-runtime/README.md", + "source_sha256": "26941ae2c64ec137c3a9dc97e8d766d93cf6bdcb103ee290f871fd0a181f2757", + "utf8_bytes": 9587, + "line_start": 127, + "line_end": 128, + "snippet": "Independent mode is never a silent fallback. The paired DAG omits it and requires\nthe preceding Astra receipt/report to match the packet and report hashes.", + "support_role": "answer_support" + } + ] + }, + { + "id": "execution-14", + "category": "deterministic_execution", + "query": "Which parts of a research report does the deterministic checker validate, and does passing it prove the prose conclusions?", + "question": "Which parts of a research report does the deterministic checker validate, and does passing it prove the prose conclusions?", + "expected_behavior": "answer_with_sources", + "expected_abstention": false, + "expected_files": [ + "blueprints/us-equities/research-runtime/README.md" + ], + "boundary_files": [], + "relevance_rationale": "The checker boundary identifies both deterministic fields and unsupported semantic or trading conclusions.", + "reference_answer": "It validates exact values, units, citation identifiers, cutoff and scope. Passing does not prove every prose inference, profitability, diligence or trading authority.", + "source_evidence": [ + { + "path": "blueprints/us-equities/research-runtime/README.md", + "source_sha256": "26941ae2c64ec137c3a9dc97e8d766d93cf6bdcb103ee290f871fd0a181f2757", + "utf8_bytes": 9587, + "line_start": 161, + "line_end": 163, + "snippet": "The report checker validates exact numerical values, units, citation IDs, cutoff\nand scope. It does not establish that every prose claim follows from its citations,\nprove profitability, perform investment diligence or authorize trading.", + "support_role": "answer_support" + } + ] + }, + { + "id": "research-15", + "category": "research_data", + "query": "When may a strategy act on a daily close-derived signal versus an intraday catalyst according to the research protocol?", + "question": "When may a strategy act on a daily close-derived signal versus an intraday catalyst according to the research protocol?", + "expected_behavior": "answer_with_sources", + "expected_abstention": false, + "expected_files": [ + "blueprints/us-equities/acceptance-wave/research-protocol.md" + ], + "boundary_files": [], + "relevance_rationale": "The timing protocol explicitly separates close-derived daily decisions from intraday information receipt and latency.", + "reference_answer": "A close-derived daily signal cannot execute at the same close. Intraday entry follows actual receipt plus declared processing/order latency when the market state is tradable.", + "source_evidence": [ + { + "path": "blueprints/us-equities/acceptance-wave/research-protocol.md", + "source_sha256": "a4bbb4b1682e84c7324b903f90ececce3a3f7c547b445517089607794ded4f92", + "utf8_bytes": 6449, + "line_start": 17, + "line_end": 19, + "snippet": "A close-derived daily signal\ncannot execute at that same close. Intraday entry follows actual information\nreceipt and declared processing/order latency at a tradable market state.", + "support_role": "answer_support" + } + ] + }, + { + "id": "research-16", + "category": "research_data", + "query": "For the extreme-mover discovery cohort, what price ratio corresponds to a 200 percent increase, and can different return horizons be pooled?", + "question": "For the extreme-mover discovery cohort, what price ratio corresponds to a 200 percent increase, and can different return horizons be pooled?", + "expected_behavior": "answer_with_sources", + "expected_abstention": false, + "expected_files": [ + "blueprints/us-equities/acceptance-wave/research-protocol.md" + ], + "boundary_files": [], + "relevance_rationale": "The cohort definition supplies the exact numerical interpretation and prevents horizon substitution.", + "reference_answer": "A return of at least 2.0 corresponds to a price ratio of at least 3.0. Freeze the anchor and horizon; those different measures must not be pooled as one return.", + "source_evidence": [ + { + "path": "blueprints/us-equities/acceptance-wave/research-protocol.md", + "source_sha256": "a4bbb4b1682e84c7324b903f90ececce3a3f7c547b445517089607794ded4f92", + "utf8_bytes": 6449, + "line_start": 24, + "line_end": 28, + "snippet": "Create a discovery cohort for observed price increases of at least 200%:\nreturn `>= 2.0`, equivalent to a price ratio `>= 3.0`. Freeze the anchor and\nhorizon before extraction. Prior regular close to intraday high, first executable\npost-signal quote to exit, and multiday close-to-close are different measures;\nnever mix them in one claimed return.", + "support_role": "answer_support" + } + ] + }, + { + "id": "research-17", + "category": "research_data", + "query": "After the reserved 2021 control segment has been scored, may a later experiment describe it as untouched held-out evidence?", + "question": "After the reserved 2021 control segment has been scored, may a later experiment describe it as untouched held-out evidence?", + "expected_behavior": "answer_with_sources", + "expected_abstention": false, + "expected_files": [ + "blueprints/us-equities/research-evaluation/README.md" + ], + "boundary_files": [], + "relevance_rationale": "The chronological control guide directly states the changed evidence status after inspection.", + "reference_answer": "No. Once scored, the segment is inspected evidence for future work and must not be described as a fresh untouched holdout.", + "source_evidence": [ + { + "path": "blueprints/us-equities/research-evaluation/README.md", + "source_sha256": "5f6e4b5cc7338e9f5fc3edeb0d2d5092f1c2ba779d0243ed3ba29da76a80edde", + "utf8_bytes": 8589, + "line_start": 34, + "line_end": 35, + "snippet": "Once scored, these results are inspected evidence\nfor future work and cannot be presented as a fresh untouched holdout.", + "support_role": "answer_support" + } + ] + }, + { + "id": "backup-18", + "category": "backup", + "query": "Does the native memory backup include model caches, external session transcripts and the raw archive as well as the application store?", + "question": "Does the native memory backup include model caches, external session transcripts and the raw archive as well as the application store?", + "expected_behavior": "answer_with_sources", + "expected_abstention": false, + "expected_files": [ + "blueprints/us-equities/state-recovery/memory/README.md" + ], + "boundary_files": [], + "relevance_rationale": "The recovery limits enumerate the backup contents and specifically excluded data classes.", + "reference_answer": "No. It includes wiki, SQLite and service configuration, but excludes raw data, model caches, logs and external native transcripts; it is not complete ledger or account recovery.", + "source_evidence": [ + { + "path": "blueprints/us-equities/state-recovery/memory/README.md", + "source_sha256": "c58fd2a181feb488281c5a0a57fe7b6227f085ad44c41c53f539632d26a34965", + "utf8_bytes": 8477, + "line_start": 141, + "line_end": 143, + "snippet": "- Native backup includes wiki, SQLite and service configuration. It excludes\n `raw/`, model caches, logs and external native session transcripts. Do not\n describe it as complete portable-ledger or account recovery.", + "support_role": "answer_support" + } + ] + }, + { + "id": "backup-19", + "category": "backup", + "query": "When verifying a recovered vector database, are matching collection and point counts enough?", + "question": "When verifying a recovered vector database, are matching collection and point counts enough?", + "expected_behavior": "answer_with_sources", + "expected_abstention": false, + "expected_files": [ + "blueprints/us-equities/state-recovery/qdrant/README.md" + ], + "boundary_files": [], + "relevance_rationale": "The vector-state restore workflow requires full logical-record and configuration comparison beyond aggregate counts.", + "reference_answer": "No. Compare canonical complete point IDs, payloads and vectors, plus full collection configuration, payload schema and aliases; counts alone are insufficient.", + "source_evidence": [ + { + "path": "blueprints/us-equities/state-recovery/qdrant/README.md", + "source_sha256": "13f5ce8292fcec69e3daa85859c800aecb4e529bf1f94391dc883129623febe0", + "utf8_bytes": 5252, + "line_start": 66, + "line_end": 69, + "snippet": "Use native `GET /collections`, `GET /collections/{name}` and\n`POST /collections/{name}/points/scroll` with payloads and vectors to compare\nlogical state; canonicalize complete point records, not just counts. Also compare\nthe complete `config`, `payload_schema` and `GET /aliases` responses.", + "support_role": "answer_support" + } + ] + }, + { + "id": "execution-20", + "category": "deterministic_execution", + "query": "If the corporate-action amount matches the reference but the provider omits the currency, should the comparison fill in USD or declare the amount inconsistent?", + "question": "If the corporate-action amount matches the reference but the provider omits the currency, should the comparison fill in USD or declare the amount inconsistent?", + "expected_behavior": "answer_with_sources", + "expected_abstention": false, + "expected_files": [ + "blueprints/us-equities/authenticated-data/README.md" + ], + "boundary_files": [], + "relevance_rationale": "The authenticated-data comparison explicitly preserves an unknown unit instead of inventing or contradicting it.", + "reference_answer": "Neither. Preserve currency as unknown and leave cash-unit reconciliation open; do not infer USD or treat the matching amount as a conflict.", + "source_evidence": [ + { + "path": "blueprints/us-equities/authenticated-data/README.md", + "source_sha256": "76d2439e456aacca301013d1993bde75e42fd6ffa0f77990eb2e60686a961666", + "utf8_bytes": 8763, + "line_start": 27, + "line_end": 29, + "snippet": "The corporate-action\ncomparison treats missing currency as unknown, rather than a conflicting amount\nor an inferred USD label.", + "support_role": "answer_support" + } + ] + }, + { + "id": "abstain-21", + "category": "native_credentials_host_evidence", + "query": "Which provider account is signed in on the next workstation we will onboard right now?", + "question": "Which provider account is signed in on the next workstation we will onboard right now?", + "expected_behavior": "abstain_from_requested_specifics", + "expected_abstention": true, + "expected_files": [], + "boundary_files": [ + "adoption/README.md" + ], + "relevance_rationale": "The public adoption guide explains why historical source-host receipts cannot reveal current authentication on an unspecified new host.", + "reference_answer": "Abstain from naming an account or asserting sign-in. The public corpus does not contain current account state for that future destination.", + "source_evidence": [ + { + "path": "adoption/README.md", + "source_sha256": "5e554f816905070b474aaa8e177fbc7b4ed81e017019fb3067ec08ca67b6c568", + "utf8_bytes": 9783, + "line_start": 3, + "line_end": 3, + "snippet": "Historical receipts describe the authoring host; a clone does not inherit its logins, service state, tool discovery or acceptance.", + "support_role": "abstention_boundary" + } + ] + }, + { + "id": "abstain-22", + "category": "context_usage", + "query": "What is the exact net billed token reduction for a complete comparable task after enabling this stack?", + "question": "What is the exact net billed token reduction for a complete comparable task after enabling this stack?", + "expected_behavior": "abstain_from_requested_specifics", + "expected_abstention": true, + "expected_files": [], + "boundary_files": [ + "blueprints/us-equities/authenticated-data/README.md" + ], + "relevance_rationale": "The current convergence guide explicitly leaves provider savings unmeasured; no exact billed reduction can be supplied.", + "reference_answer": "Abstain from giving a numeric savings figure. The corpus has no matched whole-task billed-token comparison supporting one.", + "source_evidence": [ + { + "path": "blueprints/us-equities/authenticated-data/README.md", + "source_sha256": "76d2439e456aacca301013d1993bde75e42fd6ffa0f77990eb2e60686a961666", + "utf8_bytes": 8763, + "line_start": 132, + "line_end": 132, + "snippet": "Provider tokens saved and strategy profitability are not measured.", + "support_role": "abstention_boundary" + } + ] + }, + { + "id": "abstain-23", + "category": "research_data", + "query": "Give the final numerical capital allocation, risk limits and promotion thresholds approved for deploying the catalyst strategy.", + "question": "Give the final numerical capital allocation, risk limits and promotion thresholds approved for deploying the catalyst strategy.", + "expected_behavior": "abstain_from_requested_specifics", + "expected_abstention": true, + "expected_files": [], + "boundary_files": [ + "blueprints/us-equities/acceptance-wave/research-protocol.md" + ], + "relevance_rationale": "The selected research protocol expressly leaves these choices unset and provides no deployment authorization.", + "reference_answer": "Abstain from inventing allocations, risk thresholds or authorization. Those parameters are unset and the document is a research design.", + "source_evidence": [ + { + "path": "blueprints/us-equities/acceptance-wave/research-protocol.md", + "source_sha256": "a4bbb4b1682e84c7324b903f90ececce3a3f7c547b445517089607794ded4f92", + "utf8_bytes": 6449, + "line_start": 7, + "line_end": 8, + "snippet": "leaving capital, risk, source entitlements, evaluation intervals and promotion\nthresholds unset. This is a research design, with no accepted strategy or broker authority.", + "support_role": "abstention_boundary" + } + ] + }, + { + "id": "abstain-24", + "category": "research_data", + "query": "Provide the complete historical census of all securities that rose at least 200 percent in the target research universe.", + "question": "Provide the complete historical census of all securities that rose at least 200 percent in the target research universe.", + "expected_behavior": "abstain_from_requested_specifics", + "expected_abstention": true, + "expected_files": [], + "boundary_files": [ + "blueprints/us-equities/acceptance-wave/research-protocol.md" + ], + "relevance_rationale": "The protocol identifies the requested census as data not acquired by this research wave; a complete answer is unavailable in the frozen corpus.", + "reference_answer": "Abstain from fabricating a list or claiming completeness. The frozen protocol supplies a cohort design, not an acquired complete mover census.", + "source_evidence": [ + { + "path": "blueprints/us-equities/acceptance-wave/research-protocol.md", + "source_sha256": "a4bbb4b1682e84c7324b903f90ececce3a3f7c547b445517089607794ded4f92", + "utf8_bytes": 6449, + "line_start": 100, + "line_end": 100, + "snippet": "No historical vendor dataset or +200% case census was acquired in this wave.", + "support_role": "abstention_boundary" + } + ] + } + ], + "hashfreeze": { + "algorithm": "sha256", + "canonicalization": "UTF-8 JSON, sorted keys, compact separators, ensure_ascii=false", + "payload_scope": "entire top-level object excluding hashfreeze", + "payload_sha256": "206b14cbc6ecbbafce4d5afad9d249fb767d3f4e49142d3e4020599e17808454", + "frozen_before_first_search_or_ranking_run": true, + "change_policy": "Never edit scored questions or labels in place; create a new version and retain prior results." + } +} diff --git a/blueprints/convergence-practice/local-fixture/freeze-receipt.json b/blueprints/convergence-practice/local-fixture/freeze-receipt.json new file mode 100644 index 000000000..dc21c105f --- /dev/null +++ b/blueprints/convergence-practice/local-fixture/freeze-receipt.json @@ -0,0 +1,14 @@ +{ + "created_at": "2026-09-20T02:06:03.632262+00:00", + "base_commit": "bf99d340f798bf72cc3104619edc482882eb423d", + "fixture_sha256": "aea5e9616fb19eb1d2fd58ecfadf88ff087d7d2e79b4a70bc9f88207af9d0402", + "corpus_manifest_sha256": "7bb986160b332e205208b6767b95b6f019423ffed1f50843534ff7e4272fdded", + "fixture_payload_sha256": "206b14cbc6ecbbafce4d5afad9d249fb767d3f4e49142d3e4020599e17808454", + "corpus_payload_sha256": "f104af0112d6c6307a5bc7e08e3f9e65835e9efef35b4e69c098f948058bae83", + "question_count": 24, + "answerable_count": 20, + "abstention_count": 4, + "corpus_document_count": 15, + "source_byte_total": 105003, + "no_search_or_ranking_executed": true +} diff --git a/blueprints/convergence-practice/local-fixture/rankings.json b/blueprints/convergence-practice/local-fixture/rankings.json new file mode 100644 index 000000000..07549823d --- /dev/null +++ b/blueprints/convergence-practice/local-fixture/rankings.json @@ -0,0 +1,1475 @@ +{ + "base_commit": "bf99d340f798bf72cc3104619edc482882eb423d", + "chunks": 15, + "corpus_manifest_sha256": "7bb986160b332e205208b6767b95b6f019423ffed1f50843534ff7e4272fdded", + "documents": 15, + "fixture_sha256": "aea5e9616fb19eb1d2fd58ecfadf88ff087d7d2e79b4a70bc9f88207af9d0402", + "full_body_normalization": "Native CRLF/CR to LF and surrounding whitespace strip only.", + "max_chunk_chars": 9777, + "query_transform": "None: original fixture query string passed verbatim to native ranker.", + "results": [ + { + "expected_abstention": false, + "gold_files": [ + "adoption/README.md" + ], + "metrics": { + "MRR": 1.0, + "Recall@1": 1.0, + "Recall@3": 1.0, + "Recall@5": 1.0 + }, + "ranked_files": [ + "adoption/README.md", + "blueprints/us-equities/authenticated-data/README.md", + "adoption/update.md", + "docs/evidence.md", + "blueprints/us-equities/research-runtime/README.md", + "blueprints/us-equities/state-recovery/memory/README.md", + "observability/desktop-restart.md", + "blueprints/us-equities/worker-supervision/README.md", + "blueprints/us-equities/state-recovery/README.md", + "blueprints/us-equities/research-evaluation/README.md", + "observability/session-e2e.md", + "blueprints/us-equities/state-recovery/qdrant/README.md", + "blueprints/us-equities/memory-lifecycle/README.md", + "blueprints/us-equities/acceptance-wave/research-protocol.md", + "blueprints/us-equities/data-readiness/README.md" + ], + "ranker": "lexical", + "sample_id": "credentials-01" + }, + { + "expected_abstention": false, + "gold_files": [ + "adoption/README.md" + ], + "metrics": { + "MRR": 1.0, + "Recall@1": 1.0, + "Recall@3": 1.0, + "Recall@5": 1.0 + }, + "ranked_files": [ + "adoption/README.md", + "blueprints/us-equities/worker-supervision/README.md", + "blueprints/us-equities/data-readiness/README.md", + "blueprints/us-equities/research-runtime/README.md", + "blueprints/us-equities/authenticated-data/README.md", + "blueprints/us-equities/memory-lifecycle/README.md", + "adoption/update.md", + "observability/desktop-restart.md", + "blueprints/us-equities/state-recovery/qdrant/README.md", + "docs/evidence.md", + "blueprints/us-equities/state-recovery/memory/README.md", + "blueprints/us-equities/acceptance-wave/research-protocol.md", + "observability/session-e2e.md", + "blueprints/us-equities/state-recovery/README.md", + "blueprints/us-equities/research-evaluation/README.md" + ], + "ranker": "lexical", + "sample_id": "credentials-02" + }, + { + "expected_abstention": false, + "gold_files": [ + "docs/evidence.md" + ], + "metrics": { + "MRR": 1.0, + "Recall@1": 1.0, + "Recall@3": 1.0, + "Recall@5": 1.0 + }, + "ranked_files": [ + "docs/evidence.md", + "observability/session-e2e.md", + "blueprints/us-equities/data-readiness/README.md", + "blueprints/us-equities/authenticated-data/README.md", + "blueprints/us-equities/research-runtime/README.md", + "blueprints/us-equities/research-evaluation/README.md", + "blueprints/us-equities/state-recovery/memory/README.md", + "adoption/README.md", + "adoption/update.md", + "observability/desktop-restart.md", + "blueprints/us-equities/worker-supervision/README.md", + "blueprints/us-equities/state-recovery/qdrant/README.md", + "blueprints/us-equities/memory-lifecycle/README.md", + "blueprints/us-equities/acceptance-wave/research-protocol.md", + "blueprints/us-equities/state-recovery/README.md" + ], + "ranker": "lexical", + "sample_id": "evidence-03" + }, + { + "expected_abstention": false, + "gold_files": [ + "blueprints/us-equities/memory-lifecycle/README.md" + ], + "metrics": { + "MRR": 1.0, + "Recall@1": 1.0, + "Recall@3": 1.0, + "Recall@5": 1.0 + }, + "ranked_files": [ + "blueprints/us-equities/memory-lifecycle/README.md", + "blueprints/us-equities/state-recovery/memory/README.md", + "blueprints/us-equities/worker-supervision/README.md", + "blueprints/us-equities/research-evaluation/README.md", + "docs/evidence.md", + "blueprints/us-equities/state-recovery/README.md", + "adoption/update.md", + "adoption/README.md", + "blueprints/us-equities/acceptance-wave/research-protocol.md", + "blueprints/us-equities/research-runtime/README.md", + "blueprints/us-equities/authenticated-data/README.md", + "observability/desktop-restart.md", + "blueprints/us-equities/data-readiness/README.md", + "blueprints/us-equities/state-recovery/qdrant/README.md", + "observability/session-e2e.md" + ], + "ranker": "lexical", + "sample_id": "memory-04" + }, + { + "expected_abstention": false, + "gold_files": [ + "blueprints/us-equities/memory-lifecycle/README.md" + ], + "metrics": { + "MRR": 1.0, + "Recall@1": 1.0, + "Recall@3": 1.0, + "Recall@5": 1.0 + }, + "ranked_files": [ + "blueprints/us-equities/memory-lifecycle/README.md", + "adoption/update.md", + "docs/evidence.md", + "observability/session-e2e.md", + "blueprints/us-equities/state-recovery/memory/README.md", + "adoption/README.md", + "blueprints/us-equities/authenticated-data/README.md", + "blueprints/us-equities/worker-supervision/README.md", + "blueprints/us-equities/state-recovery/qdrant/README.md", + "blueprints/us-equities/data-readiness/README.md", + "blueprints/us-equities/acceptance-wave/research-protocol.md", + "observability/desktop-restart.md", + "blueprints/us-equities/research-runtime/README.md", + "blueprints/us-equities/state-recovery/README.md", + "blueprints/us-equities/research-evaluation/README.md" + ], + "ranker": "lexical", + "sample_id": "memory-05" + }, + { + "expected_abstention": false, + "gold_files": [ + "blueprints/us-equities/memory-lifecycle/README.md" + ], + "metrics": { + "MRR": 1.0, + "Recall@1": 1.0, + "Recall@3": 1.0, + "Recall@5": 1.0 + }, + "ranked_files": [ + "blueprints/us-equities/memory-lifecycle/README.md", + "docs/evidence.md", + "observability/desktop-restart.md", + "adoption/update.md", + "blueprints/us-equities/authenticated-data/README.md", + "blueprints/us-equities/worker-supervision/README.md", + "blueprints/us-equities/data-readiness/README.md", + "blueprints/us-equities/acceptance-wave/research-protocol.md", + "blueprints/us-equities/state-recovery/memory/README.md", + "observability/session-e2e.md", + "blueprints/us-equities/research-evaluation/README.md", + "blueprints/us-equities/state-recovery/qdrant/README.md", + "adoption/README.md", + "blueprints/us-equities/state-recovery/README.md", + "blueprints/us-equities/research-runtime/README.md" + ], + "ranker": "lexical", + "sample_id": "memory-06" + }, + { + "expected_abstention": false, + "gold_files": [ + "observability/desktop-restart.md" + ], + "metrics": { + "MRR": 0.5, + "Recall@1": 0.0, + "Recall@3": 1.0, + "Recall@5": 1.0 + }, + "ranked_files": [ + "adoption/README.md", + "observability/desktop-restart.md", + "blueprints/us-equities/research-runtime/README.md", + "adoption/update.md", + "docs/evidence.md", + "blueprints/us-equities/worker-supervision/README.md", + "blueprints/us-equities/memory-lifecycle/README.md", + "observability/session-e2e.md", + "blueprints/us-equities/authenticated-data/README.md", + "blueprints/us-equities/state-recovery/memory/README.md", + "blueprints/us-equities/research-evaluation/README.md", + "blueprints/us-equities/data-readiness/README.md", + "blueprints/us-equities/acceptance-wave/research-protocol.md", + "blueprints/us-equities/state-recovery/README.md", + "blueprints/us-equities/state-recovery/qdrant/README.md" + ], + "ranker": "lexical", + "sample_id": "context-07" + }, + { + "expected_abstention": false, + "gold_files": [ + "observability/session-e2e.md" + ], + "metrics": { + "MRR": 0.5, + "Recall@1": 0.0, + "Recall@3": 1.0, + "Recall@5": 1.0 + }, + "ranked_files": [ + "adoption/update.md", + "observability/session-e2e.md", + "observability/desktop-restart.md", + "blueprints/us-equities/research-evaluation/README.md", + "blueprints/us-equities/acceptance-wave/research-protocol.md", + "blueprints/us-equities/research-runtime/README.md", + "blueprints/us-equities/data-readiness/README.md", + "blueprints/us-equities/state-recovery/qdrant/README.md", + "blueprints/us-equities/authenticated-data/README.md", + "blueprints/us-equities/worker-supervision/README.md", + "docs/evidence.md", + "blueprints/us-equities/memory-lifecycle/README.md", + "adoption/README.md", + "blueprints/us-equities/state-recovery/README.md", + "blueprints/us-equities/state-recovery/memory/README.md" + ], + "ranker": "lexical", + "sample_id": "context-08" + }, + { + "expected_abstention": false, + "gold_files": [ + "observability/desktop-restart.md" + ], + "metrics": { + "MRR": 1.0, + "Recall@1": 1.0, + "Recall@3": 1.0, + "Recall@5": 1.0 + }, + "ranked_files": [ + "observability/desktop-restart.md", + "docs/evidence.md", + "blueprints/us-equities/research-runtime/README.md", + "observability/session-e2e.md", + "adoption/README.md", + "blueprints/us-equities/worker-supervision/README.md", + "blueprints/us-equities/memory-lifecycle/README.md", + "adoption/update.md", + "blueprints/us-equities/authenticated-data/README.md", + "blueprints/us-equities/data-readiness/README.md", + "blueprints/us-equities/state-recovery/memory/README.md", + "blueprints/us-equities/state-recovery/qdrant/README.md", + "blueprints/us-equities/research-evaluation/README.md", + "blueprints/us-equities/acceptance-wave/research-protocol.md", + "blueprints/us-equities/state-recovery/README.md" + ], + "ranker": "lexical", + "sample_id": "context-09" + }, + { + "expected_abstention": false, + "gold_files": [ + "blueprints/us-equities/worker-supervision/README.md" + ], + "metrics": { + "MRR": 0.5, + "Recall@1": 0.0, + "Recall@3": 1.0, + "Recall@5": 1.0 + }, + "ranked_files": [ + "blueprints/us-equities/research-runtime/README.md", + "blueprints/us-equities/worker-supervision/README.md", + "blueprints/us-equities/authenticated-data/README.md", + "blueprints/us-equities/data-readiness/README.md", + "blueprints/us-equities/state-recovery/qdrant/README.md", + "docs/evidence.md", + "adoption/README.md", + "blueprints/us-equities/state-recovery/memory/README.md", + "adoption/update.md", + "blueprints/us-equities/research-evaluation/README.md", + "observability/session-e2e.md", + "blueprints/us-equities/state-recovery/README.md", + "observability/desktop-restart.md", + "blueprints/us-equities/acceptance-wave/research-protocol.md", + "blueprints/us-equities/memory-lifecycle/README.md" + ], + "ranker": "lexical", + "sample_id": "worker-10" + }, + { + "expected_abstention": false, + "gold_files": [ + "blueprints/us-equities/worker-supervision/README.md" + ], + "metrics": { + "MRR": 1.0, + "Recall@1": 1.0, + "Recall@3": 1.0, + "Recall@5": 1.0 + }, + "ranked_files": [ + "blueprints/us-equities/worker-supervision/README.md", + "observability/session-e2e.md", + "blueprints/us-equities/state-recovery/README.md", + "blueprints/us-equities/memory-lifecycle/README.md", + "blueprints/us-equities/state-recovery/memory/README.md", + "blueprints/us-equities/state-recovery/qdrant/README.md", + "adoption/update.md", + "blueprints/us-equities/authenticated-data/README.md", + "blueprints/us-equities/research-runtime/README.md", + "adoption/README.md", + "docs/evidence.md", + "observability/desktop-restart.md", + "blueprints/us-equities/acceptance-wave/research-protocol.md", + "blueprints/us-equities/research-evaluation/README.md", + "blueprints/us-equities/data-readiness/README.md" + ], + "ranker": "lexical", + "sample_id": "worker-11" + }, + { + "expected_abstention": false, + "gold_files": [ + "blueprints/us-equities/worker-supervision/README.md" + ], + "metrics": { + "MRR": 1.0, + "Recall@1": 1.0, + "Recall@3": 1.0, + "Recall@5": 1.0 + }, + "ranked_files": [ + "blueprints/us-equities/worker-supervision/README.md", + "blueprints/us-equities/state-recovery/memory/README.md", + "blueprints/us-equities/state-recovery/qdrant/README.md", + "observability/session-e2e.md", + "docs/evidence.md", + "blueprints/us-equities/data-readiness/README.md", + "blueprints/us-equities/state-recovery/README.md", + "blueprints/us-equities/research-runtime/README.md", + "adoption/README.md", + "blueprints/us-equities/authenticated-data/README.md", + "observability/desktop-restart.md", + "adoption/update.md", + "blueprints/us-equities/memory-lifecycle/README.md", + "blueprints/us-equities/acceptance-wave/research-protocol.md", + "blueprints/us-equities/research-evaluation/README.md" + ], + "ranker": "lexical", + "sample_id": "worker-12" + }, + { + "expected_abstention": false, + "gold_files": [ + "blueprints/us-equities/research-runtime/README.md" + ], + "metrics": { + "MRR": 0.5, + "Recall@1": 0.0, + "Recall@3": 1.0, + "Recall@5": 1.0 + }, + "ranked_files": [ + "adoption/README.md", + "blueprints/us-equities/research-runtime/README.md", + "observability/desktop-restart.md", + "blueprints/us-equities/authenticated-data/README.md", + "adoption/update.md", + "docs/evidence.md", + "blueprints/us-equities/state-recovery/README.md", + "blueprints/us-equities/data-readiness/README.md", + "observability/session-e2e.md", + "blueprints/us-equities/acceptance-wave/research-protocol.md", + "blueprints/us-equities/worker-supervision/README.md", + "blueprints/us-equities/memory-lifecycle/README.md", + "blueprints/us-equities/research-evaluation/README.md", + "blueprints/us-equities/state-recovery/memory/README.md", + "blueprints/us-equities/state-recovery/qdrant/README.md" + ], + "ranker": "lexical", + "sample_id": "execution-13" + }, + { + "expected_abstention": false, + "gold_files": [ + "blueprints/us-equities/research-runtime/README.md" + ], + "metrics": { + "MRR": 1.0, + "Recall@1": 1.0, + "Recall@3": 1.0, + "Recall@5": 1.0 + }, + "ranked_files": [ + "blueprints/us-equities/research-runtime/README.md", + "adoption/README.md", + "adoption/update.md", + "blueprints/us-equities/worker-supervision/README.md", + "blueprints/us-equities/acceptance-wave/research-protocol.md", + "observability/session-e2e.md", + "blueprints/us-equities/authenticated-data/README.md", + "blueprints/us-equities/state-recovery/memory/README.md", + "blueprints/us-equities/state-recovery/qdrant/README.md", + "blueprints/us-equities/memory-lifecycle/README.md", + "blueprints/us-equities/research-evaluation/README.md", + "blueprints/us-equities/data-readiness/README.md", + "docs/evidence.md", + "blueprints/us-equities/state-recovery/README.md", + "observability/desktop-restart.md" + ], + "ranker": "lexical", + "sample_id": "execution-14" + }, + { + "expected_abstention": false, + "gold_files": [ + "blueprints/us-equities/acceptance-wave/research-protocol.md" + ], + "metrics": { + "MRR": 1.0, + "Recall@1": 1.0, + "Recall@3": 1.0, + "Recall@5": 1.0 + }, + "ranked_files": [ + "blueprints/us-equities/acceptance-wave/research-protocol.md", + "blueprints/us-equities/authenticated-data/README.md", + "adoption/update.md", + "blueprints/us-equities/data-readiness/README.md", + "blueprints/us-equities/research-runtime/README.md", + "blueprints/us-equities/state-recovery/README.md", + "adoption/README.md", + "blueprints/us-equities/research-evaluation/README.md", + "observability/session-e2e.md", + "blueprints/us-equities/state-recovery/memory/README.md", + "blueprints/us-equities/memory-lifecycle/README.md", + "docs/evidence.md", + "observability/desktop-restart.md", + "blueprints/us-equities/worker-supervision/README.md", + "blueprints/us-equities/state-recovery/qdrant/README.md" + ], + "ranker": "lexical", + "sample_id": "research-15" + }, + { + "expected_abstention": false, + "gold_files": [ + "blueprints/us-equities/acceptance-wave/research-protocol.md" + ], + "metrics": { + "MRR": 1.0, + "Recall@1": 1.0, + "Recall@3": 1.0, + "Recall@5": 1.0 + }, + "ranked_files": [ + "blueprints/us-equities/acceptance-wave/research-protocol.md", + "blueprints/us-equities/data-readiness/README.md", + "adoption/update.md", + "blueprints/us-equities/authenticated-data/README.md", + "blueprints/us-equities/research-evaluation/README.md", + "observability/desktop-restart.md", + "blueprints/us-equities/worker-supervision/README.md", + "docs/evidence.md", + "blueprints/us-equities/state-recovery/memory/README.md", + "adoption/README.md", + "observability/session-e2e.md", + "blueprints/us-equities/memory-lifecycle/README.md", + "blueprints/us-equities/state-recovery/README.md", + "blueprints/us-equities/research-runtime/README.md", + "blueprints/us-equities/state-recovery/qdrant/README.md" + ], + "ranker": "lexical", + "sample_id": "research-16" + }, + { + "expected_abstention": false, + "gold_files": [ + "blueprints/us-equities/research-evaluation/README.md" + ], + "metrics": { + "MRR": 1.0, + "Recall@1": 1.0, + "Recall@3": 1.0, + "Recall@5": 1.0 + }, + "ranked_files": [ + "blueprints/us-equities/research-evaluation/README.md", + "blueprints/us-equities/acceptance-wave/research-protocol.md", + "blueprints/us-equities/worker-supervision/README.md", + "blueprints/us-equities/state-recovery/memory/README.md", + "blueprints/us-equities/authenticated-data/README.md", + "adoption/README.md", + "observability/session-e2e.md", + "blueprints/us-equities/memory-lifecycle/README.md", + "adoption/update.md", + "blueprints/us-equities/research-runtime/README.md", + "observability/desktop-restart.md", + "docs/evidence.md", + "blueprints/us-equities/data-readiness/README.md", + "blueprints/us-equities/state-recovery/qdrant/README.md", + "blueprints/us-equities/state-recovery/README.md" + ], + "ranker": "lexical", + "sample_id": "research-17" + }, + { + "expected_abstention": false, + "gold_files": [ + "blueprints/us-equities/state-recovery/memory/README.md" + ], + "metrics": { + "MRR": 1.0, + "Recall@1": 1.0, + "Recall@3": 1.0, + "Recall@5": 1.0 + }, + "ranked_files": [ + "blueprints/us-equities/state-recovery/memory/README.md", + "blueprints/us-equities/state-recovery/README.md", + "blueprints/us-equities/memory-lifecycle/README.md", + "blueprints/us-equities/worker-supervision/README.md", + "adoption/README.md", + "docs/evidence.md", + "observability/session-e2e.md", + "blueprints/us-equities/authenticated-data/README.md", + "adoption/update.md", + "blueprints/us-equities/research-runtime/README.md", + "blueprints/us-equities/acceptance-wave/research-protocol.md", + "blueprints/us-equities/research-evaluation/README.md", + "blueprints/us-equities/data-readiness/README.md", + "blueprints/us-equities/state-recovery/qdrant/README.md", + "observability/desktop-restart.md" + ], + "ranker": "lexical", + "sample_id": "backup-18" + }, + { + "expected_abstention": false, + "gold_files": [ + "blueprints/us-equities/state-recovery/qdrant/README.md" + ], + "metrics": { + "MRR": 1.0, + "Recall@1": 1.0, + "Recall@3": 1.0, + "Recall@5": 1.0 + }, + "ranked_files": [ + "blueprints/us-equities/state-recovery/qdrant/README.md", + "blueprints/us-equities/state-recovery/README.md", + "adoption/update.md", + "observability/desktop-restart.md", + "docs/evidence.md", + "blueprints/us-equities/state-recovery/memory/README.md", + "adoption/README.md", + "blueprints/us-equities/research-runtime/README.md", + "blueprints/us-equities/research-evaluation/README.md", + "blueprints/us-equities/memory-lifecycle/README.md", + "blueprints/us-equities/data-readiness/README.md", + "observability/session-e2e.md", + "blueprints/us-equities/authenticated-data/README.md", + "blueprints/us-equities/worker-supervision/README.md", + "blueprints/us-equities/acceptance-wave/research-protocol.md" + ], + "ranker": "lexical", + "sample_id": "backup-19" + }, + { + "expected_abstention": false, + "gold_files": [ + "blueprints/us-equities/authenticated-data/README.md" + ], + "metrics": { + "MRR": 1.0, + "Recall@1": 1.0, + "Recall@3": 1.0, + "Recall@5": 1.0 + }, + "ranked_files": [ + "blueprints/us-equities/authenticated-data/README.md", + "blueprints/us-equities/data-readiness/README.md", + "blueprints/us-equities/research-evaluation/README.md", + "adoption/update.md", + "blueprints/us-equities/research-runtime/README.md", + "blueprints/us-equities/memory-lifecycle/README.md", + "blueprints/us-equities/state-recovery/memory/README.md", + "adoption/README.md", + "blueprints/us-equities/state-recovery/README.md", + "observability/desktop-restart.md", + "observability/session-e2e.md", + "blueprints/us-equities/acceptance-wave/research-protocol.md", + "blueprints/us-equities/state-recovery/qdrant/README.md", + "docs/evidence.md", + "blueprints/us-equities/worker-supervision/README.md" + ], + "ranker": "lexical", + "sample_id": "execution-20" + }, + { + "expected_abstention": true, + "gold_files": [], + "metrics": null, + "ranked_files": [ + "blueprints/us-equities/authenticated-data/README.md", + "observability/session-e2e.md", + "blueprints/us-equities/research-runtime/README.md", + "blueprints/us-equities/memory-lifecycle/README.md", + "observability/desktop-restart.md", + "adoption/README.md", + "adoption/update.md", + "blueprints/us-equities/acceptance-wave/research-protocol.md", + "blueprints/us-equities/data-readiness/README.md", + "blueprints/us-equities/worker-supervision/README.md", + "blueprints/us-equities/research-evaluation/README.md", + "blueprints/us-equities/state-recovery/qdrant/README.md", + "blueprints/us-equities/state-recovery/README.md", + "docs/evidence.md", + "blueprints/us-equities/state-recovery/memory/README.md" + ], + "ranker": "lexical", + "sample_id": "abstain-21" + }, + { + "expected_abstention": true, + "gold_files": [], + "metrics": null, + "ranked_files": [ + "docs/evidence.md", + "observability/session-e2e.md", + "blueprints/us-equities/state-recovery/qdrant/README.md", + "blueprints/us-equities/authenticated-data/README.md", + "adoption/update.md", + "observability/desktop-restart.md", + "blueprints/us-equities/state-recovery/memory/README.md", + "blueprints/us-equities/research-runtime/README.md", + "adoption/README.md", + "blueprints/us-equities/worker-supervision/README.md", + "blueprints/us-equities/research-evaluation/README.md", + "blueprints/us-equities/state-recovery/README.md", + "blueprints/us-equities/memory-lifecycle/README.md", + "blueprints/us-equities/acceptance-wave/research-protocol.md", + "blueprints/us-equities/data-readiness/README.md" + ], + "ranker": "lexical", + "sample_id": "abstain-22" + }, + { + "expected_abstention": true, + "gold_files": [], + "metrics": null, + "ranked_files": [ + "blueprints/us-equities/acceptance-wave/research-protocol.md", + "blueprints/us-equities/research-runtime/README.md", + "blueprints/us-equities/data-readiness/README.md", + "adoption/update.md", + "blueprints/us-equities/authenticated-data/README.md", + "observability/session-e2e.md", + "blueprints/us-equities/state-recovery/README.md", + "observability/desktop-restart.md", + "blueprints/us-equities/memory-lifecycle/README.md", + "blueprints/us-equities/worker-supervision/README.md", + "docs/evidence.md", + "blueprints/us-equities/research-evaluation/README.md", + "blueprints/us-equities/state-recovery/memory/README.md", + "blueprints/us-equities/state-recovery/qdrant/README.md", + "adoption/README.md" + ], + "ranker": "lexical", + "sample_id": "abstain-23" + }, + { + "expected_abstention": true, + "gold_files": [], + "metrics": null, + "ranked_files": [ + "blueprints/us-equities/acceptance-wave/research-protocol.md", + "blueprints/us-equities/authenticated-data/README.md", + "blueprints/us-equities/state-recovery/memory/README.md", + "adoption/update.md", + "blueprints/us-equities/state-recovery/qdrant/README.md", + "blueprints/us-equities/research-evaluation/README.md", + "blueprints/us-equities/data-readiness/README.md", + "observability/desktop-restart.md", + "adoption/README.md", + "observability/session-e2e.md", + "docs/evidence.md", + "blueprints/us-equities/memory-lifecycle/README.md", + "blueprints/us-equities/research-runtime/README.md", + "blueprints/us-equities/worker-supervision/README.md", + "blueprints/us-equities/state-recovery/README.md" + ], + "ranker": "lexical", + "sample_id": "abstain-24" + }, + { + "expected_abstention": false, + "gold_files": [ + "adoption/README.md" + ], + "metrics": { + "MRR": 1.0, + "Recall@1": 1.0, + "Recall@3": 1.0, + "Recall@5": 1.0 + }, + "ranked_files": [ + "adoption/README.md", + "docs/evidence.md", + "adoption/update.md", + "blueprints/us-equities/authenticated-data/README.md", + "blueprints/us-equities/research-runtime/README.md", + "observability/desktop-restart.md", + "observability/session-e2e.md", + "blueprints/us-equities/state-recovery/memory/README.md", + "blueprints/us-equities/research-evaluation/README.md", + "blueprints/us-equities/worker-supervision/README.md", + "blueprints/us-equities/state-recovery/README.md", + "blueprints/us-equities/acceptance-wave/research-protocol.md", + "blueprints/us-equities/memory-lifecycle/README.md", + "blueprints/us-equities/state-recovery/qdrant/README.md", + "blueprints/us-equities/data-readiness/README.md" + ], + "ranker": "bm25", + "sample_id": "credentials-01" + }, + { + "expected_abstention": false, + "gold_files": [ + "adoption/README.md" + ], + "metrics": { + "MRR": 1.0, + "Recall@1": 1.0, + "Recall@3": 1.0, + "Recall@5": 1.0 + }, + "ranked_files": [ + "adoption/README.md", + "blueprints/us-equities/research-runtime/README.md", + "blueprints/us-equities/data-readiness/README.md", + "blueprints/us-equities/authenticated-data/README.md", + "blueprints/us-equities/worker-supervision/README.md", + "docs/evidence.md", + "blueprints/us-equities/acceptance-wave/research-protocol.md", + "adoption/update.md", + "blueprints/us-equities/memory-lifecycle/README.md", + "observability/desktop-restart.md", + "observability/session-e2e.md", + "blueprints/us-equities/state-recovery/memory/README.md", + "blueprints/us-equities/state-recovery/qdrant/README.md", + "blueprints/us-equities/research-evaluation/README.md", + "blueprints/us-equities/state-recovery/README.md" + ], + "ranker": "bm25", + "sample_id": "credentials-02" + }, + { + "expected_abstention": false, + "gold_files": [ + "docs/evidence.md" + ], + "metrics": { + "MRR": 1.0, + "Recall@1": 1.0, + "Recall@3": 1.0, + "Recall@5": 1.0 + }, + "ranked_files": [ + "docs/evidence.md", + "observability/session-e2e.md", + "blueprints/us-equities/authenticated-data/README.md", + "adoption/README.md", + "blueprints/us-equities/data-readiness/README.md", + "adoption/update.md", + "blueprints/us-equities/state-recovery/memory/README.md", + "blueprints/us-equities/research-evaluation/README.md", + "blueprints/us-equities/research-runtime/README.md", + "observability/desktop-restart.md", + "blueprints/us-equities/worker-supervision/README.md", + "blueprints/us-equities/acceptance-wave/research-protocol.md", + "blueprints/us-equities/memory-lifecycle/README.md", + "blueprints/us-equities/state-recovery/qdrant/README.md", + "blueprints/us-equities/state-recovery/README.md" + ], + "ranker": "bm25", + "sample_id": "evidence-03" + }, + { + "expected_abstention": false, + "gold_files": [ + "blueprints/us-equities/memory-lifecycle/README.md" + ], + "metrics": { + "MRR": 1.0, + "Recall@1": 1.0, + "Recall@3": 1.0, + "Recall@5": 1.0 + }, + "ranked_files": [ + "blueprints/us-equities/memory-lifecycle/README.md", + "blueprints/us-equities/state-recovery/memory/README.md", + "blueprints/us-equities/research-evaluation/README.md", + "blueprints/us-equities/acceptance-wave/research-protocol.md", + "adoption/README.md", + "blueprints/us-equities/worker-supervision/README.md", + "docs/evidence.md", + "blueprints/us-equities/research-runtime/README.md", + "blueprints/us-equities/authenticated-data/README.md", + "adoption/update.md", + "observability/session-e2e.md", + "blueprints/us-equities/data-readiness/README.md", + "observability/desktop-restart.md", + "blueprints/us-equities/state-recovery/qdrant/README.md", + "blueprints/us-equities/state-recovery/README.md" + ], + "ranker": "bm25", + "sample_id": "memory-04" + }, + { + "expected_abstention": false, + "gold_files": [ + "blueprints/us-equities/memory-lifecycle/README.md" + ], + "metrics": { + "MRR": 1.0, + "Recall@1": 1.0, + "Recall@3": 1.0, + "Recall@5": 1.0 + }, + "ranked_files": [ + "blueprints/us-equities/memory-lifecycle/README.md", + "docs/evidence.md", + "observability/session-e2e.md", + "adoption/update.md", + "blueprints/us-equities/worker-supervision/README.md", + "adoption/README.md", + "blueprints/us-equities/state-recovery/memory/README.md", + "blueprints/us-equities/authenticated-data/README.md", + "blueprints/us-equities/acceptance-wave/research-protocol.md", + "blueprints/us-equities/state-recovery/qdrant/README.md", + "blueprints/us-equities/research-runtime/README.md", + "observability/desktop-restart.md", + "blueprints/us-equities/data-readiness/README.md", + "blueprints/us-equities/state-recovery/README.md", + "blueprints/us-equities/research-evaluation/README.md" + ], + "ranker": "bm25", + "sample_id": "memory-05" + }, + { + "expected_abstention": false, + "gold_files": [ + "blueprints/us-equities/memory-lifecycle/README.md" + ], + "metrics": { + "MRR": 1.0, + "Recall@1": 1.0, + "Recall@3": 1.0, + "Recall@5": 1.0 + }, + "ranked_files": [ + "blueprints/us-equities/memory-lifecycle/README.md", + "docs/evidence.md", + "observability/desktop-restart.md", + "blueprints/us-equities/authenticated-data/README.md", + "adoption/update.md", + "blueprints/us-equities/acceptance-wave/research-protocol.md", + "blueprints/us-equities/data-readiness/README.md", + "blueprints/us-equities/state-recovery/memory/README.md", + "blueprints/us-equities/worker-supervision/README.md", + "observability/session-e2e.md", + "blueprints/us-equities/state-recovery/qdrant/README.md", + "adoption/README.md", + "blueprints/us-equities/research-evaluation/README.md", + "blueprints/us-equities/research-runtime/README.md", + "blueprints/us-equities/state-recovery/README.md" + ], + "ranker": "bm25", + "sample_id": "memory-06" + }, + { + "expected_abstention": false, + "gold_files": [ + "observability/desktop-restart.md" + ], + "metrics": { + "MRR": 1.0, + "Recall@1": 1.0, + "Recall@3": 1.0, + "Recall@5": 1.0 + }, + "ranked_files": [ + "observability/desktop-restart.md", + "adoption/README.md", + "blueprints/us-equities/research-runtime/README.md", + "docs/evidence.md", + "adoption/update.md", + "blueprints/us-equities/worker-supervision/README.md", + "observability/session-e2e.md", + "blueprints/us-equities/memory-lifecycle/README.md", + "blueprints/us-equities/authenticated-data/README.md", + "blueprints/us-equities/state-recovery/memory/README.md", + "blueprints/us-equities/research-evaluation/README.md", + "blueprints/us-equities/acceptance-wave/research-protocol.md", + "blueprints/us-equities/data-readiness/README.md", + "blueprints/us-equities/state-recovery/qdrant/README.md", + "blueprints/us-equities/state-recovery/README.md" + ], + "ranker": "bm25", + "sample_id": "context-07" + }, + { + "expected_abstention": false, + "gold_files": [ + "observability/session-e2e.md" + ], + "metrics": { + "MRR": 1.0, + "Recall@1": 1.0, + "Recall@3": 1.0, + "Recall@5": 1.0 + }, + "ranked_files": [ + "observability/session-e2e.md", + "adoption/update.md", + "blueprints/us-equities/research-evaluation/README.md", + "blueprints/us-equities/acceptance-wave/research-protocol.md", + "observability/desktop-restart.md", + "blueprints/us-equities/authenticated-data/README.md", + "blueprints/us-equities/research-runtime/README.md", + "adoption/README.md", + "docs/evidence.md", + "blueprints/us-equities/data-readiness/README.md", + "blueprints/us-equities/worker-supervision/README.md", + "blueprints/us-equities/memory-lifecycle/README.md", + "blueprints/us-equities/state-recovery/qdrant/README.md", + "blueprints/us-equities/state-recovery/memory/README.md", + "blueprints/us-equities/state-recovery/README.md" + ], + "ranker": "bm25", + "sample_id": "context-08" + }, + { + "expected_abstention": false, + "gold_files": [ + "observability/desktop-restart.md" + ], + "metrics": { + "MRR": 1.0, + "Recall@1": 1.0, + "Recall@3": 1.0, + "Recall@5": 1.0 + }, + "ranked_files": [ + "observability/desktop-restart.md", + "docs/evidence.md", + "observability/session-e2e.md", + "blueprints/us-equities/research-runtime/README.md", + "adoption/README.md", + "blueprints/us-equities/worker-supervision/README.md", + "adoption/update.md", + "blueprints/us-equities/authenticated-data/README.md", + "blueprints/us-equities/memory-lifecycle/README.md", + "blueprints/us-equities/state-recovery/memory/README.md", + "blueprints/us-equities/state-recovery/qdrant/README.md", + "blueprints/us-equities/data-readiness/README.md", + "blueprints/us-equities/research-evaluation/README.md", + "blueprints/us-equities/acceptance-wave/research-protocol.md", + "blueprints/us-equities/state-recovery/README.md" + ], + "ranker": "bm25", + "sample_id": "context-09" + }, + { + "expected_abstention": false, + "gold_files": [ + "blueprints/us-equities/worker-supervision/README.md" + ], + "metrics": { + "MRR": 0.3333333333333333, + "Recall@1": 0.0, + "Recall@3": 1.0, + "Recall@5": 1.0 + }, + "ranked_files": [ + "blueprints/us-equities/research-runtime/README.md", + "blueprints/us-equities/authenticated-data/README.md", + "blueprints/us-equities/worker-supervision/README.md", + "blueprints/us-equities/data-readiness/README.md", + "docs/evidence.md", + "adoption/README.md", + "adoption/update.md", + "observability/session-e2e.md", + "blueprints/us-equities/state-recovery/memory/README.md", + "blueprints/us-equities/research-evaluation/README.md", + "blueprints/us-equities/state-recovery/qdrant/README.md", + "blueprints/us-equities/acceptance-wave/research-protocol.md", + "observability/desktop-restart.md", + "blueprints/us-equities/state-recovery/README.md", + "blueprints/us-equities/memory-lifecycle/README.md" + ], + "ranker": "bm25", + "sample_id": "worker-10" + }, + { + "expected_abstention": false, + "gold_files": [ + "blueprints/us-equities/worker-supervision/README.md" + ], + "metrics": { + "MRR": 1.0, + "Recall@1": 1.0, + "Recall@3": 1.0, + "Recall@5": 1.0 + }, + "ranked_files": [ + "blueprints/us-equities/worker-supervision/README.md", + "observability/session-e2e.md", + "blueprints/us-equities/state-recovery/memory/README.md", + "adoption/README.md", + "blueprints/us-equities/state-recovery/README.md", + "blueprints/us-equities/memory-lifecycle/README.md", + "blueprints/us-equities/state-recovery/qdrant/README.md", + "adoption/update.md", + "blueprints/us-equities/authenticated-data/README.md", + "docs/evidence.md", + "blueprints/us-equities/research-runtime/README.md", + "observability/desktop-restart.md", + "blueprints/us-equities/acceptance-wave/research-protocol.md", + "blueprints/us-equities/research-evaluation/README.md", + "blueprints/us-equities/data-readiness/README.md" + ], + "ranker": "bm25", + "sample_id": "worker-11" + }, + { + "expected_abstention": false, + "gold_files": [ + "blueprints/us-equities/worker-supervision/README.md" + ], + "metrics": { + "MRR": 1.0, + "Recall@1": 1.0, + "Recall@3": 1.0, + "Recall@5": 1.0 + }, + "ranked_files": [ + "blueprints/us-equities/worker-supervision/README.md", + "docs/evidence.md", + "blueprints/us-equities/state-recovery/memory/README.md", + "blueprints/us-equities/state-recovery/qdrant/README.md", + "observability/session-e2e.md", + "adoption/README.md", + "blueprints/us-equities/data-readiness/README.md", + "blueprints/us-equities/research-runtime/README.md", + "blueprints/us-equities/authenticated-data/README.md", + "adoption/update.md", + "observability/desktop-restart.md", + "blueprints/us-equities/memory-lifecycle/README.md", + "blueprints/us-equities/state-recovery/README.md", + "blueprints/us-equities/research-evaluation/README.md", + "blueprints/us-equities/acceptance-wave/research-protocol.md" + ], + "ranker": "bm25", + "sample_id": "worker-12" + }, + { + "expected_abstention": false, + "gold_files": [ + "blueprints/us-equities/research-runtime/README.md" + ], + "metrics": { + "MRR": 1.0, + "Recall@1": 1.0, + "Recall@3": 1.0, + "Recall@5": 1.0 + }, + "ranked_files": [ + "blueprints/us-equities/research-runtime/README.md", + "adoption/README.md", + "blueprints/us-equities/authenticated-data/README.md", + "docs/evidence.md", + "adoption/update.md", + "observability/session-e2e.md", + "observability/desktop-restart.md", + "blueprints/us-equities/data-readiness/README.md", + "blueprints/us-equities/acceptance-wave/research-protocol.md", + "blueprints/us-equities/state-recovery/README.md", + "blueprints/us-equities/worker-supervision/README.md", + "blueprints/us-equities/memory-lifecycle/README.md", + "blueprints/us-equities/research-evaluation/README.md", + "blueprints/us-equities/state-recovery/memory/README.md", + "blueprints/us-equities/state-recovery/qdrant/README.md" + ], + "ranker": "bm25", + "sample_id": "execution-13" + }, + { + "expected_abstention": false, + "gold_files": [ + "blueprints/us-equities/research-runtime/README.md" + ], + "metrics": { + "MRR": 1.0, + "Recall@1": 1.0, + "Recall@3": 1.0, + "Recall@5": 1.0 + }, + "ranked_files": [ + "blueprints/us-equities/research-runtime/README.md", + "adoption/README.md", + "adoption/update.md", + "blueprints/us-equities/authenticated-data/README.md", + "blueprints/us-equities/acceptance-wave/research-protocol.md", + "observability/session-e2e.md", + "blueprints/us-equities/worker-supervision/README.md", + "blueprints/us-equities/research-evaluation/README.md", + "blueprints/us-equities/state-recovery/memory/README.md", + "blueprints/us-equities/state-recovery/qdrant/README.md", + "blueprints/us-equities/memory-lifecycle/README.md", + "docs/evidence.md", + "blueprints/us-equities/data-readiness/README.md", + "observability/desktop-restart.md", + "blueprints/us-equities/state-recovery/README.md" + ], + "ranker": "bm25", + "sample_id": "execution-14" + }, + { + "expected_abstention": false, + "gold_files": [ + "blueprints/us-equities/acceptance-wave/research-protocol.md" + ], + "metrics": { + "MRR": 1.0, + "Recall@1": 1.0, + "Recall@3": 1.0, + "Recall@5": 1.0 + }, + "ranked_files": [ + "blueprints/us-equities/acceptance-wave/research-protocol.md", + "blueprints/us-equities/authenticated-data/README.md", + "adoption/update.md", + "blueprints/us-equities/data-readiness/README.md", + "blueprints/us-equities/research-runtime/README.md", + "adoption/README.md", + "blueprints/us-equities/research-evaluation/README.md", + "observability/session-e2e.md", + "blueprints/us-equities/state-recovery/memory/README.md", + "docs/evidence.md", + "blueprints/us-equities/memory-lifecycle/README.md", + "observability/desktop-restart.md", + "blueprints/us-equities/worker-supervision/README.md", + "blueprints/us-equities/state-recovery/qdrant/README.md", + "blueprints/us-equities/state-recovery/README.md" + ], + "ranker": "bm25", + "sample_id": "research-15" + }, + { + "expected_abstention": false, + "gold_files": [ + "blueprints/us-equities/acceptance-wave/research-protocol.md" + ], + "metrics": { + "MRR": 1.0, + "Recall@1": 1.0, + "Recall@3": 1.0, + "Recall@5": 1.0 + }, + "ranked_files": [ + "blueprints/us-equities/acceptance-wave/research-protocol.md", + "blueprints/us-equities/data-readiness/README.md", + "adoption/update.md", + "blueprints/us-equities/research-evaluation/README.md", + "blueprints/us-equities/authenticated-data/README.md", + "observability/desktop-restart.md", + "docs/evidence.md", + "blueprints/us-equities/worker-supervision/README.md", + "observability/session-e2e.md", + "adoption/README.md", + "blueprints/us-equities/state-recovery/memory/README.md", + "blueprints/us-equities/memory-lifecycle/README.md", + "blueprints/us-equities/research-runtime/README.md", + "blueprints/us-equities/state-recovery/qdrant/README.md", + "blueprints/us-equities/state-recovery/README.md" + ], + "ranker": "bm25", + "sample_id": "research-16" + }, + { + "expected_abstention": false, + "gold_files": [ + "blueprints/us-equities/research-evaluation/README.md" + ], + "metrics": { + "MRR": 1.0, + "Recall@1": 1.0, + "Recall@3": 1.0, + "Recall@5": 1.0 + }, + "ranked_files": [ + "blueprints/us-equities/research-evaluation/README.md", + "blueprints/us-equities/acceptance-wave/research-protocol.md", + "blueprints/us-equities/worker-supervision/README.md", + "blueprints/us-equities/authenticated-data/README.md", + "blueprints/us-equities/state-recovery/memory/README.md", + "adoption/README.md", + "observability/session-e2e.md", + "blueprints/us-equities/memory-lifecycle/README.md", + "adoption/update.md", + "blueprints/us-equities/research-runtime/README.md", + "observability/desktop-restart.md", + "docs/evidence.md", + "blueprints/us-equities/data-readiness/README.md", + "blueprints/us-equities/state-recovery/qdrant/README.md", + "blueprints/us-equities/state-recovery/README.md" + ], + "ranker": "bm25", + "sample_id": "research-17" + }, + { + "expected_abstention": false, + "gold_files": [ + "blueprints/us-equities/state-recovery/memory/README.md" + ], + "metrics": { + "MRR": 1.0, + "Recall@1": 1.0, + "Recall@3": 1.0, + "Recall@5": 1.0 + }, + "ranked_files": [ + "blueprints/us-equities/state-recovery/memory/README.md", + "blueprints/us-equities/state-recovery/README.md", + "blueprints/us-equities/memory-lifecycle/README.md", + "docs/evidence.md", + "adoption/README.md", + "blueprints/us-equities/worker-supervision/README.md", + "observability/session-e2e.md", + "blueprints/us-equities/authenticated-data/README.md", + "blueprints/us-equities/acceptance-wave/research-protocol.md", + "blueprints/us-equities/research-runtime/README.md", + "adoption/update.md", + "blueprints/us-equities/research-evaluation/README.md", + "blueprints/us-equities/state-recovery/qdrant/README.md", + "blueprints/us-equities/data-readiness/README.md", + "observability/desktop-restart.md" + ], + "ranker": "bm25", + "sample_id": "backup-18" + }, + { + "expected_abstention": false, + "gold_files": [ + "blueprints/us-equities/state-recovery/qdrant/README.md" + ], + "metrics": { + "MRR": 1.0, + "Recall@1": 1.0, + "Recall@3": 1.0, + "Recall@5": 1.0 + }, + "ranked_files": [ + "blueprints/us-equities/state-recovery/qdrant/README.md", + "observability/desktop-restart.md", + "adoption/update.md", + "blueprints/us-equities/state-recovery/memory/README.md", + "docs/evidence.md", + "blueprints/us-equities/state-recovery/README.md", + "adoption/README.md", + "blueprints/us-equities/research-evaluation/README.md", + "blueprints/us-equities/research-runtime/README.md", + "blueprints/us-equities/authenticated-data/README.md", + "observability/session-e2e.md", + "blueprints/us-equities/data-readiness/README.md", + "blueprints/us-equities/memory-lifecycle/README.md", + "blueprints/us-equities/worker-supervision/README.md", + "blueprints/us-equities/acceptance-wave/research-protocol.md" + ], + "ranker": "bm25", + "sample_id": "backup-19" + }, + { + "expected_abstention": false, + "gold_files": [ + "blueprints/us-equities/authenticated-data/README.md" + ], + "metrics": { + "MRR": 1.0, + "Recall@1": 1.0, + "Recall@3": 1.0, + "Recall@5": 1.0 + }, + "ranked_files": [ + "blueprints/us-equities/authenticated-data/README.md", + "blueprints/us-equities/data-readiness/README.md", + "blueprints/us-equities/research-evaluation/README.md", + "adoption/update.md", + "blueprints/us-equities/research-runtime/README.md", + "blueprints/us-equities/memory-lifecycle/README.md", + "blueprints/us-equities/state-recovery/memory/README.md", + "blueprints/us-equities/acceptance-wave/research-protocol.md", + "adoption/README.md", + "observability/desktop-restart.md", + "observability/session-e2e.md", + "docs/evidence.md", + "blueprints/us-equities/state-recovery/qdrant/README.md", + "blueprints/us-equities/state-recovery/README.md", + "blueprints/us-equities/worker-supervision/README.md" + ], + "ranker": "bm25", + "sample_id": "execution-20" + }, + { + "expected_abstention": true, + "gold_files": [], + "metrics": null, + "ranked_files": [ + "blueprints/us-equities/authenticated-data/README.md", + "observability/session-e2e.md", + "adoption/update.md", + "adoption/README.md", + "blueprints/us-equities/research-runtime/README.md", + "observability/desktop-restart.md", + "blueprints/us-equities/research-evaluation/README.md", + "blueprints/us-equities/acceptance-wave/research-protocol.md", + "docs/evidence.md", + "blueprints/us-equities/worker-supervision/README.md", + "blueprints/us-equities/memory-lifecycle/README.md", + "blueprints/us-equities/data-readiness/README.md", + "blueprints/us-equities/state-recovery/memory/README.md", + "blueprints/us-equities/state-recovery/qdrant/README.md", + "blueprints/us-equities/state-recovery/README.md" + ], + "ranker": "bm25", + "sample_id": "abstain-21" + }, + { + "expected_abstention": true, + "gold_files": [], + "metrics": null, + "ranked_files": [ + "docs/evidence.md", + "observability/session-e2e.md", + "blueprints/us-equities/authenticated-data/README.md", + "observability/desktop-restart.md", + "adoption/README.md", + "adoption/update.md", + "blueprints/us-equities/research-runtime/README.md", + "blueprints/us-equities/state-recovery/qdrant/README.md", + "blueprints/us-equities/research-evaluation/README.md", + "blueprints/us-equities/state-recovery/memory/README.md", + "blueprints/us-equities/worker-supervision/README.md", + "blueprints/us-equities/acceptance-wave/research-protocol.md", + "blueprints/us-equities/memory-lifecycle/README.md", + "blueprints/us-equities/data-readiness/README.md", + "blueprints/us-equities/state-recovery/README.md" + ], + "ranker": "bm25", + "sample_id": "abstain-22" + }, + { + "expected_abstention": true, + "gold_files": [], + "metrics": null, + "ranked_files": [ + "blueprints/us-equities/acceptance-wave/research-protocol.md", + "blueprints/us-equities/research-runtime/README.md", + "adoption/update.md", + "blueprints/us-equities/data-readiness/README.md", + "blueprints/us-equities/authenticated-data/README.md", + "observability/session-e2e.md", + "blueprints/us-equities/state-recovery/README.md", + "observability/desktop-restart.md", + "docs/evidence.md", + "blueprints/us-equities/research-evaluation/README.md", + "blueprints/us-equities/memory-lifecycle/README.md", + "blueprints/us-equities/worker-supervision/README.md", + "blueprints/us-equities/state-recovery/memory/README.md", + "adoption/README.md", + "blueprints/us-equities/state-recovery/qdrant/README.md" + ], + "ranker": "bm25", + "sample_id": "abstain-23" + }, + { + "expected_abstention": true, + "gold_files": [], + "metrics": null, + "ranked_files": [ + "blueprints/us-equities/acceptance-wave/research-protocol.md", + "blueprints/us-equities/authenticated-data/README.md", + "blueprints/us-equities/state-recovery/memory/README.md", + "blueprints/us-equities/research-evaluation/README.md", + "adoption/update.md", + "adoption/README.md", + "blueprints/us-equities/state-recovery/qdrant/README.md", + "observability/desktop-restart.md", + "docs/evidence.md", + "blueprints/us-equities/data-readiness/README.md", + "observability/session-e2e.md", + "blueprints/us-equities/memory-lifecycle/README.md", + "blueprints/us-equities/research-runtime/README.md", + "blueprints/us-equities/worker-supervision/README.md", + "blueprints/us-equities/state-recovery/README.md" + ], + "ranker": "bm25", + "sample_id": "abstain-24" + } + ], + "schema_version": 1, + "summaries": { + "bm25": { + "abstention_accuracy": null, + "abstention_policy": null, + "macro_metrics": { + "MRR": 0.9666666666666666, + "Recall@1": 0.95, + "Recall@3": 1.0, + "Recall@5": 1.0 + }, + "negative_rankings_returned": 4, + "negative_samples_excluded_from_positive_metrics": 4, + "positive_samples": 20, + "wall_seconds": 0.07384333299705759 + }, + "lexical": { + "abstention_accuracy": null, + "abstention_policy": null, + "macro_metrics": { + "MRR": 0.9, + "Recall@1": 0.8, + "Recall@3": 1.0, + "Recall@5": 1.0 + }, + "negative_rankings_returned": 4, + "negative_samples_excluded_from_positive_metrics": 4, + "positive_samples": 20, + "wall_seconds": 0.06728295799985062 + } + } +} diff --git a/blueprints/convergence-practice/protocol.json b/blueprints/convergence-practice/protocol.json new file mode 100644 index 000000000..d8569bae9 --- /dev/null +++ b/blueprints/convergence-practice/protocol.json @@ -0,0 +1,32 @@ +{ + "schema_version": 1, + "kind": "convergence_experiment_contract", + "scope": "Reusable research-to-acceptance practice; no installation, host qualification, model run or trading authority follows from inclusion.", + "evidence_classes": ["discovery", "pinned_source_review", "offline_artifact_check", "native_cli_execution", "model_task_execution", "recovery_execution"], + "adoption_decisions": ["retain", "trial", "adopt_within_scope", "defer", "reject"], + "required_experiment_fields": [ + "task_and_failure", "primary_sources_and_pins", "discovery_provenance", + "license_and_compatibility", "baseline_and_candidate", "frozen_inputs", + "predeclared_metrics", "runtime_and_platform", "commands", "observations", + "failures_and_skips", "limitations", "decision_and_scope", "rollback", + "next_decision_changing_test" + ], + "lanes": [ + {"id": "retrieval", "acceptance": "Exact-source coverage plus explicit no-answer cases; task outcomes measured separately.", "next_gate": "Held-out task evaluation on a substantial real project."}, + {"id": "native-agent-engineering", "acceptance": "A real issue produces a tested independently reviewed patch in an isolated workspace.", "next_gate": "Matched correctness, elapsed time and complete usage comparison."}, + {"id": "memory", "acceptance": "An explicit durable decision is retrievable in its intended scope with cross-scope negatives.", "next_gate": "Recovery without importing unrelated histories or another host's credential store."}, + {"id": "access-recovery", "acceptance": "Interruption, reconnect, cancellation and selected-state restoration have observed results.", "next_gate": "Independent-host restore and controlled startup/network-change qualification."}, + {"id": "local-inference", "acceptance": "A pinned model/runtime meets task quality and measured resource limits against the existing baseline.", "next_gate": "Useful native client integration; inference speed alone does not pass."}, + {"id": "research-application", "acceptance": "A reproducible source-to-result packet preserves time, provenance and deterministic calculations.", "next_gate": "Application-specific data and execution gates remain independently enforced."} + ], + "publication_policy": { + "public_sources_only": true, + "private_repository_export": false, + "credential_export": false, + "automatic_transcript_capture": false, + "automatic_install_or_promotion": false, + "historical_host_acceptance_transfers": false, + "provider_savings_from_artifact_reduction": false, + "universal_sota_claim": false + } +} diff --git a/catalogs/convergence-practice/public-owned.json b/catalogs/convergence-practice/public-owned.json new file mode 100644 index 000000000..0898fdd2c --- /dev/null +++ b/catalogs/convergence-practice/public-owned.json @@ -0,0 +1,22 @@ +{ + "schema_version": 1, + "kind": "public_owned", + "endpoint": "https://api.github.com/users/seathatflowsinourveins/repos?type=owner&per_page=100", + "retrieved_at": "2026-09-20T02:00:50.538618+00:00", + "count": 1, + "filtered_nonpublic_count": 0, + "fields_allowlisted": true, + "repositories": [ + { + "repository": "seathatflowsinourveins/native-agent-stack", + "url": "https://github.com/seathatflowsinourveins/native-agent-stack", + "description": "Evidence-backed native Codex and Claude stack: scoped memory, automatic local code RAG, context efficiency, upstream recipes and reproducible receipts.", + "default_branch": "main", + "archived": false, + "fork": false, + "pushed_at": "2026-09-20T01:20:58Z", + "license_metadata": "MIT", + "topics": [] + } + ] +} diff --git a/catalogs/convergence-practice/public-starred.json b/catalogs/convergence-practice/public-starred.json new file mode 100644 index 000000000..e31bb7fad --- /dev/null +++ b/catalogs/convergence-practice/public-starred.json @@ -0,0 +1,7398 @@ +{ + "schema_version": 1, + "kind": "public_starred", + "endpoint": "https://api.github.com/users/seathatflowsinourveins/starred?per_page=100", + "retrieved_at": "2026-09-20T02:00:50.538618+00:00", + "count": 342, + "filtered_nonpublic_count": 0, + "fields_allowlisted": true, + "repositories": [ + { + "repository": "Storybloq/storybloq", + "url": "https://github.com/Storybloq/storybloq", + "description": "Project memory and workflows for Claude Code and Codex. Keep stories, plans, handovers, and review evidence in your repo. Resume across sessions and follow progress in the Mac app.", + "default_branch": "main", + "archived": false, + "fork": false, + "pushed_at": "2026-09-20T00:59:15Z", + "license_metadata": "NOASSERTION", + "topics": [ + "agentic-development", + "ai-development", + "anthropic", + "claude-code", + "claude-skill", + "cli", + "context-management", + "developer-tools", + "handover", + "mac-app", + "macos", + "mcp", + "mcp-server", + "project-management", + "session-continuity", + "typescript", + "workflow" + ] + }, + { + "repository": "Nexting-ai/nexting", + "url": "https://github.com/Nexting-ai/nexting", + "description": "Remote control for Claude Code, Codex, Grok, and Cursor on Mac or PC. View sessions, send tasks, and drive them remotely from your phone, PIN, or Ring. OpenClaw supported.", + "default_branch": "main", + "archived": false, + "fork": false, + "pushed_at": "2026-09-11T03:12:02Z", + "license_metadata": "MIT", + "topics": [ + "agent-remote-control", + "ai-agent", + "ai-wearable", + "claude-code", + "codex", + "hardware", + "nrf52840", + "openclaw", + "remote-control", + "wearable" + ] + }, + { + "repository": "tw93/Mole", + "url": "https://github.com/tw93/Mole", + "description": "\ud83d\udc39 Clean, uninstall, analyze, optimize, and monitor your Mac. Free open-source CLI, plus a native Mac app.", + "default_branch": "main", + "archived": false, + "fork": false, + "pushed_at": "2026-09-20T01:58:45Z", + "license_metadata": "GPL-3.0", + "topics": [ + "analyzer", + "appcleaner", + "clean", + "cleaner", + "cleanmymac", + "command-line", + "daisydisk", + "istat", + "mac", + "macos", + "macos-app", + "native", + "optimize", + "pearcleaner", + "sensei", + "shell", + "swift", + "swiftui", + "uninstall" + ] + }, + { + "repository": "clash-verge-rev/clash-verge-rev", + "url": "https://github.com/clash-verge-rev/clash-verge-rev", + "description": "A modern GUI client based on Tauri, designed to run in Windows, macOS and Linux for tailored proxy experience", + "default_branch": "dev", + "archived": false, + "fork": false, + "pushed_at": "2026-09-20T01:23:42Z", + "license_metadata": "GPL-3.0", + "topics": [ + "clash", + "clash-meta", + "clash-verge", + "linux", + "mac", + "mihomo", + "tauri-app", + "windows" + ] + }, + { + "repository": "jaywcjlove/awesome-mac", + "url": "https://github.com/jaywcjlove/awesome-mac", + "description": "\uf8ff This project is dedicated to collecting high-quality macOS software and organizing them systematically by different categories for easy search and use.", + "default_branch": "master", + "archived": false, + "fork": false, + "pushed_at": "2026-09-17T08:01:12Z", + "license_metadata": "CC0-1.0", + "topics": [ + "app", + "apple", + "application", + "apps", + "awesome", + "awesome-list", + "awesome-mac", + "desktop-app", + "desktop-application", + "desktop-apps", + "list", + "mac", + "mac-osx", + "macos", + "macos-app", + "macos-apps", + "macosx", + "software", + "swift", + "swiftui" + ] + }, + { + "repository": "enaqx/awesome-react", + "url": "https://github.com/enaqx/awesome-react", + "description": "A collection of awesome things regarding React ecosystem", + "default_branch": "master", + "archived": false, + "fork": false, + "pushed_at": "2026-09-04T01:13:20Z", + "license_metadata": null, + "topics": [ + "awesome", + "awesome-list", + "javascript", + "react", + "react-apps", + "react-native", + "react-tutorial", + "samples", + "tutorial", + "typescript" + ] + }, + { + "repository": "tech-leads-club/agent-skills", + "url": "https://github.com/tech-leads-club/agent-skills", + "description": "The secure, validated skill registry for professional AI coding agents. Extend Antigravity, Claude Code, Cursor, Copilot and more with absolute confidence.", + "default_branch": "main", + "archived": false, + "fork": false, + "pushed_at": "2026-09-19T14:45:51Z", + "license_metadata": "NOASSERTION", + "topics": [ + "agent", + "ai", + "antigravity", + "claude-code", + "copilot", + "cursor", + "skills" + ] + }, + { + "repository": "Tencent/WeKnora", + "url": "https://github.com/Tencent/WeKnora", + "description": "Open-source LLM knowledge platform: turn raw documents into a queryable RAG, an autonomous reasoning agent, and a self-maintaining Wiki.", + "default_branch": "main", + "archived": false, + "fork": false, + "pushed_at": "2026-09-19T16:17:11Z", + "license_metadata": "NOASSERTION", + "topics": [ + "agent", + "agentic", + "ai", + "chatbot", + "dsh-plugin", + "embeddings", + "evaluation", + "generative-ai", + "golang", + "knowledge-base", + "llm", + "multi-tenant", + "ollama", + "openai", + "question-answering", + "rag", + "reranking", + "semantic-search", + "vector-search", + "wiki" + ] + }, + { + "repository": "n8n-io/n8n", + "url": "https://github.com/n8n-io/n8n", + "description": "Fair-code workflow automation platform with native AI capabilities. Combine visual building with custom code, self-host or cloud, 400+ integrations.", + "default_branch": "master", + "archived": false, + "fork": false, + "pushed_at": "2026-09-19T20:30:25Z", + "license_metadata": "NOASSERTION", + "topics": [ + "ai", + "apis", + "automation", + "cli", + "data-flow", + "development", + "integration-framework", + "integrations", + "ipaas", + "low-code", + "low-code-platform", + "mcp", + "mcp-client", + "mcp-server", + "n8n", + "no-code", + "self-hosted", + "typescript", + "workflow", + "workflow-automation" + ] + }, + { + "repository": "cline/cline", + "url": "https://github.com/cline/cline", + "description": "Autonomous coding agent as an SDK, IDE extension, or CLI assistant.", + "default_branch": "main", + "archived": false, + "fork": false, + "pushed_at": "2026-09-19T21:55:43Z", + "license_metadata": "Apache-2.0", + "topics": [] + }, + { + "repository": "alphaXiv/OpenResearch", + "url": "https://github.com/alphaXiv/OpenResearch", + "description": "Turn your coding agents into research agents", + "default_branch": "main", + "archived": false, + "fork": false, + "pushed_at": "2026-09-19T05:26:17Z", + "license_metadata": "MIT", + "topics": [] + }, + { + "repository": "alibaba/open-code-review", + "url": "https://github.com/alibaba/open-code-review", + "description": "Secure, fast, efficient, battle-tested at Alibaba's scale. Hybrid architecture code review tool: deterministic pipelines + LLM Agent, precise line-level comments, built-in multi-language ruleset (NPE, thread-safety, XSS, SQL injection), OpenAI & Anthropic compatible.", + "default_branch": "main", + "archived": false, + "fork": false, + "pushed_at": "2026-09-19T12:30:32Z", + "license_metadata": "Apache-2.0", + "topics": [ + "agent", + "agent-skills", + "code-review", + "code-review-assistant", + "harness", + "repository-level-context" + ] + }, + { + "repository": "toon-format/toon", + "url": "https://github.com/toon-format/toon", + "description": "\ud83c\udf92 Token-Oriented Object Notation (TOON) \u2013 compact, human-readable serialization of JSON data for LLM prompts. TypeScript SDK, CLI, benchmarks.", + "default_branch": "main", + "archived": false, + "fork": false, + "pushed_at": "2026-09-03T07:09:24Z", + "license_metadata": "MIT", + "topics": [ + "data-format", + "llm", + "serialization", + "tokenization" + ] + }, + { + "repository": "xbtlin/ai-berkshire", + "url": "https://github.com/xbtlin/ai-berkshire", + "description": "AI \u65f6\u4ee3\u7684\u4f2f\u514b\u5e0c\u5c14\uff1a\u57fa\u4e8e Claude Code / Codex \u7684\u4ef7\u503c\u6295\u8d44\u7814\u7a76\u6846\u67b6\u3002\u5df4\u83f2\u7279\u00b7\u8292\u683c\u00b7\u6bb5\u6c38\u5e73\u00b7\u674e\u5f55\u56db\u5927\u5e08\u65b9\u6cd5\u8bba + \u591aAgent\u5e76\u884c\u7814\u7a76\u3002| AI-era Berkshire: a value investing research framework built for Claude Code / Codex. 4 masters' methodologies + multi-agent adversarial analysis.", + "default_branch": "main", + "archived": false, + "fork": false, + "pushed_at": "2026-09-19T17:42:32Z", + "license_metadata": "MIT", + "topics": [ + "ai", + "ai-agent", + "anthropic", + "berkshire-hathaway", + "charlie-munger", + "china-stock", + "claude", + "claude-code", + "financial-analysis", + "fintech", + "fundamental-analysis", + "investment", + "investment-research", + "llm", + "mcp", + "portfolio-management", + "stock-analysis", + "stock-market", + "value-investing", + "warren-buffett" + ] + }, + { + "repository": "screenpipe/screenpipe", + "url": "https://github.com/screenpipe/screenpipe", + "description": "YC (S26) | Open Computer History | Record your screen continuously locally and provide context to your agents (Claude, Codex, Openclaw, Hermes, Runner...)", + "default_branch": "main", + "archived": false, + "fork": false, + "pushed_at": "2026-09-20T01:05:19Z", + "license_metadata": "NOASSERTION", + "topics": [ + "agents", + "agi", + "ai", + "ai-memory", + "audio-recording", + "computer-vision", + "hermes", + "hermes-agent", + "llm", + "local-ai", + "local-first", + "machine-learning", + "mcp", + "multimodal", + "openclaw", + "privacy", + "rewind", + "screen-recording", + "speech-to-text", + "ycombinator" + ] + }, + { + "repository": "Leonxlnx/taste-skill", + "url": "https://github.com/Leonxlnx/taste-skill", + "description": "Taste-Skill - gives your AI good taste. stops the AI from generating boring, generic slop ", + "default_branch": "main", + "archived": false, + "fork": false, + "pushed_at": "2026-09-16T15:01:43Z", + "license_metadata": "MIT", + "topics": [ + "agent", + "ai", + "claude", + "claude-code", + "codex", + "coding", + "design", + "frontend", + "lowcode", + "nocode", + "skill", + "skills", + "vibecoding" + ] + }, + { + "repository": "nextlevelbuilder/ui-ux-pro-max-skill", + "url": "https://github.com/nextlevelbuilder/ui-ux-pro-max-skill", + "description": "An AI skill that provides design intelligence for building professional UI/UX across multiple platforms.", + "default_branch": "main", + "archived": false, + "fork": false, + "pushed_at": "2026-09-19T00:58:38Z", + "license_metadata": "MIT", + "topics": [ + "ai-skills", + "antigravity", + "claude", + "claude-code", + "codex", + "command-line", + "copilot", + "cursor-ai", + "html5", + "kiro", + "landing-page", + "mobile-ui", + "qoder", + "react", + "tailwindcss", + "trae", + "ui-design", + "uikit", + "windsurf-ai" + ] + }, + { + "repository": "langgenius/dify", + "url": "https://github.com/langgenius/dify", + "description": "Build Agentic workflows, RAG pipelines, with rich AI model and tool support on one collaborative workspace. Deploy on cloud, VPC, or self-hosted, so teams move from prototype to production without rebuilding the stack.", + "default_branch": "main", + "archived": false, + "fork": false, + "pushed_at": "2026-09-20T01:59:24Z", + "license_metadata": "NOASSERTION", + "topics": [ + "agent", + "agentic-ai", + "agentic-framework", + "agentic-workflow", + "ai", + "automation", + "claude", + "deepseek", + "genai", + "gpt", + "llm", + "low-code", + "mcp", + "nextjs", + "no-code", + "openai", + "python", + "skills", + "workflow" + ] + }, + { + "repository": "magnitudedev/magnitude", + "url": "https://github.com/magnitudedev/magnitude", + "description": "Open source inference engine optimized for consumer hardware. Profiles your machine, recommends the best models for it, then downloads, tunes, and runs them. Works on Apple Silicon, NVIDIA, AMD, or nothing but a CPU.", + "default_branch": "main", + "archived": false, + "fork": false, + "pushed_at": "2026-09-20T01:41:57Z", + "license_metadata": "Apache-2.0", + "topics": [] + }, + { + "repository": "fmtlib/fmt", + "url": "https://github.com/fmtlib/fmt", + "description": "A modern formatting library", + "default_branch": "main", + "archived": false, + "fork": false, + "pushed_at": "2026-09-19T14:49:16Z", + "license_metadata": "MIT", + "topics": [ + "c-plus-plus", + "chrono", + "cpp", + "cross-platform", + "floating-point", + "formatting", + "multiplatform", + "output", + "performance", + "printf", + "ranges", + "unicode" + ] + }, + { + "repository": "loopx-project/loopx", + "url": "https://github.com/loopx-project/loopx", + "description": "Long-horizon agent control plane for durable, governed work across Codex, Claude Code, and other harnesses.", + "default_branch": "main", + "archived": false, + "fork": false, + "pushed_at": "2026-09-20T01:58:32Z", + "license_metadata": "Apache-2.0", + "topics": [ + "agent-control-plane", + "agent-harness", + "agent-ops", + "ai-agents", + "codex", + "dsh-plugin", + "long-horizon", + "long-horizon-agents", + "loop-engineering", + "loopx", + "workflow-automation" + ] + }, + { + "repository": "paperswithbacktest/awesome-systematic-trading", + "url": "https://github.com/paperswithbacktest/awesome-systematic-trading", + "description": "A curated list of awesome libraries, packages, strategies, books, blogs, tutorials for systematic trading.", + "default_branch": "main", + "archived": false, + "fork": false, + "pushed_at": "2026-09-03T09:13:55Z", + "license_metadata": null, + "topics": [ + "algorithmic-trading", + "algotrading", + "alpha", + "arbitrage-bot", + "awesome", + "awesome-list", + "book", + "finance", + "futures", + "futures-historical-data", + "futures-market", + "futuresmarkets", + "paper", + "quant", + "quantitative-finance", + "quantitative-trading", + "trading-algorithms", + "trading-bot", + "trading-strategies" + ] + }, + { + "repository": "fmzquant/strategies", + "url": "https://github.com/fmzquant/strategies", + "description": "quantitative trading with Javascript, Python, C++, PineScript, Blockly, MyLanguage(\u9ea6\u8bed\u8a00)", + "default_branch": "master", + "archived": false, + "fork": false, + "pushed_at": "2025-04-30T04:08:14Z", + "license_metadata": null, + "topics": [] + }, + { + "repository": "bbfamily/abu", + "url": "https://github.com/bbfamily/abu", + "description": "\u963f\u5e03\u91cf\u5316\u4ea4\u6613\u7cfb\u7edf(\u80a1\u7968\uff0c\u671f\u6743\uff0c\u671f\u8d27\uff0c\u6bd4\u7279\u5e01\uff0c\u673a\u5668\u5b66\u4e60) \u57fa\u4e8epython\u7684\u5f00\u6e90\u91cf\u5316\u4ea4\u6613\uff0c\u91cf\u5316\u6295\u8d44\u67b6\u6784", + "default_branch": "master", + "archived": false, + "fork": false, + "pushed_at": "2026-01-24T09:00:29Z", + "license_metadata": "GPL-3.0", + "topics": [ + "algorithmic-trading", + "bitcoin", + "machine-learning", + "matplotlib", + "numpy", + "pandas", + "quant", + "quantitative-trading", + "stock", + "trade", + "trading" + ] + }, + { + "repository": "asavinov/intelligent-trading-bot", + "url": "https://github.com/asavinov/intelligent-trading-bot", + "description": "Intelligent Trading Bot: Automatically generating signals and trading based on machine learning and feature engineering", + "default_branch": "master", + "archived": false, + "fork": false, + "pushed_at": "2026-08-30T13:11:26Z", + "license_metadata": "MIT", + "topics": [ + "algorithmic-trading", + "artificial-intelligence", + "bitcoin", + "crypto", + "crypto-trading", + "cryptocurrency", + "feature-engineering", + "machine-learning", + "trading", + "trading-bots" + ] + }, + { + "repository": "kernc/backtesting.py", + "url": "https://github.com/kernc/backtesting.py", + "description": "\ud83d\udd0e \ud83d\udcc8 \ud83d\udc0d \ud83d\udcb0 Backtest trading strategies in Python.", + "default_branch": "master", + "archived": false, + "fork": false, + "pushed_at": "2026-08-05T12:39:16Z", + "license_metadata": "AGPL-3.0", + "topics": [ + "algo-trading", + "algorithmic-trading", + "backtesting", + "backtesting-engine", + "backtesting-frameworks", + "backtesting-trading-strategies", + "finance", + "financial-markets", + "forex", + "forex-trading", + "framework", + "hacktoberfest", + "investing", + "investment", + "investment-strategies", + "stocks", + "trading", + "trading-algorithms", + "trading-simulator", + "trading-strategies" + ] + }, + { + "repository": "whchien/ai-trader", + "url": "https://github.com/whchien/ai-trader", + "description": "Backtrader-powered backtesting framework for algorithmic trading, featuring 20+ strategies, multi-market support, CLI tools, and an integrated MCP server for professional traders.", + "default_branch": "main", + "archived": false, + "fork": false, + "pushed_at": "2026-03-28T11:16:18Z", + "license_metadata": "GPL-3.0", + "topics": [ + "ai", + "ai-agents", + "algorithmic-trading", + "backtrader", + "cli", + "mcp", + "mcp-server", + "portfolio-management", + "quantitative-finance", + "stock-price-prediction", + "trading" + ] + }, + { + "repository": "jdx/mise", + "url": "https://github.com/jdx/mise", + "description": "dev tools, env vars, task runner", + "default_branch": "main", + "archived": false, + "fork": false, + "pushed_at": "2026-09-20T01:53:19Z", + "license_metadata": "MIT", + "topics": [] + }, + { + "repository": "pola-rs/polars", + "url": "https://github.com/pola-rs/polars", + "description": "Extremely fast Query Engine for DataFrames, written in Rust", + "default_branch": "main", + "archived": false, + "fork": false, + "pushed_at": "2026-09-19T19:00:39Z", + "license_metadata": "MIT", + "topics": [ + "arrow", + "dataframe", + "dataframe-library", + "dataframes", + "out-of-core", + "polars", + "python", + "rust" + ] + }, + { + "repository": "warp-tech/warpgate", + "url": "https://github.com/warp-tech/warpgate", + "description": "Fully transparent SSH, HTTPS, Kubernetes, database and RDP/VNC bastion/PAM that doesn't need additional client-side software", + "default_branch": "main", + "archived": false, + "fork": false, + "pushed_at": "2026-09-19T20:08:00Z", + "license_metadata": "Apache-2.0", + "topics": [ + "bastion", + "bastion-host", + "https", + "https-proxy", + "infrastructure", + "kubectl", + "kubernetes", + "mysql", + "mysql-proxy", + "pam", + "postgresql-proxy", + "privileged-access-management", + "proxy", + "rdp", + "rdp-access", + "rdp-gateway", + "rust", + "ssh", + "ssh-server", + "vnc-server" + ] + }, + { + "repository": "ZSeven-W/openpencil", + "url": "https://github.com/ZSeven-W/openpencil", + "description": "The world's first open-source AI-native vector design tool and the first to feature concurrent Agent Teams. Design-as-Code. Turn prompts into UI directly on the live canvas. A modern alternative to Pencil.", + "default_branch": "main", + "archived": false, + "fork": false, + "pushed_at": "2026-09-17T15:34:46Z", + "license_metadata": "MIT", + "topics": [ + "agent", + "agent-team", + "ai", + "claude", + "claude-code", + "codex", + "dsh-plugin", + "fimga", + "flutter", + "mcp", + "opencode", + "openpencil", + "pencil", + "react", + "react-native", + "rust", + "skill", + "ui", + "vibecoding", + "vibedesign" + ] + }, + { + "repository": "eugr/spark-vllm-docker", + "url": "https://github.com/eugr/spark-vllm-docker", + "description": "Docker configuration for running VLLM on dual DGX Sparks", + "default_branch": "main", + "archived": false, + "fork": false, + "pushed_at": "2026-09-18T22:40:32Z", + "license_metadata": "MIT", + "topics": [] + }, + { + "repository": "llm-d/llm-d", + "url": "https://github.com/llm-d/llm-d", + "description": "Achieve state of the art inference performance with modern accelerators on Kubernetes", + "default_branch": "main", + "archived": false, + "fork": false, + "pushed_at": "2026-09-19T13:13:01Z", + "license_metadata": "Apache-2.0", + "topics": [ + "ai", + "cncf", + "distributed-inference", + "gpu", + "inference", + "intelligent-routing", + "kubernetes", + "llm", + "model-server" + ] + }, + { + "repository": "gastownhall/beads", + "url": "https://github.com/gastownhall/beads", + "description": "Beads - A memory upgrade for your coding agent", + "default_branch": "main", + "archived": false, + "fork": false, + "pushed_at": "2026-09-19T21:37:19Z", + "license_metadata": "MIT", + "topics": [ + "agents", + "claude-code", + "coding" + ] + }, + { + "repository": "superplanehq/superplane", + "url": "https://github.com/superplanehq/superplane", + "description": "Open source factory for one-shot engineering", + "default_branch": "main", + "archived": false, + "fork": false, + "pushed_at": "2026-09-19T03:19:55Z", + "license_metadata": "Apache-2.0", + "topics": [ + "automation", + "control-plane", + "devops", + "event-driven", + "go", + "kubernetes", + "platform-engineering", + "react", + "release-automation", + "self-hosted", + "workflow-automation" + ] + }, + { + "repository": "gitleaks/gitleaks", + "url": "https://github.com/gitleaks/gitleaks", + "description": "Find secrets with Gitleaks \ud83d\udd11", + "default_branch": "master", + "archived": false, + "fork": false, + "pushed_at": "2026-09-09T04:18:49Z", + "license_metadata": "MIT", + "topics": [ + "ai-powered", + "ci-cd", + "cicd", + "cli", + "data-loss-prevention", + "devsecops", + "dlp", + "git", + "gitleaks", + "go", + "golang", + "hacktoberfest", + "llm", + "llm-inference", + "llm-training", + "nhi", + "open-source", + "secret", + "security", + "security-tools" + ] + }, + { + "repository": "infiniflow/ragflow", + "url": "https://github.com/infiniflow/ragflow", + "description": "RAGFlow is a leading open-source Retrieval-Augmented Generation (RAG) engine that fuses cutting-edge RAG with Agent capabilities to create a superior context layer for LLMs", + "default_branch": "main", + "archived": false, + "fork": false, + "pushed_at": "2026-09-20T01:55:54Z", + "license_metadata": "Apache-2.0", + "topics": [ + "agent-harness", + "agentic-ai", + "agentic-nagive", + "agentic-retrieval", + "agentic-search", + "ai", + "ai-agents", + "context-engine", + "context-engineering", + "context-management", + "harness-engineering", + "knowledge-compilation", + "rag", + "retrieval-augmented-generation", + "search-harness" + ] + }, + { + "repository": "ollama/ollama", + "url": "https://github.com/ollama/ollama", + "description": "Get up and running with Kimi, GLM, MiniMax, DeepSeek, gpt-oss, Qwen, Gemma and other models.", + "default_branch": "main", + "archived": false, + "fork": false, + "pushed_at": "2026-09-19T20:41:40Z", + "license_metadata": "MIT", + "topics": [ + "deepseek", + "gemma", + "gemma3", + "glm", + "go", + "golang", + "gpt-oss", + "llama", + "llama3", + "llm", + "llms", + "minimax", + "mistral", + "ollama", + "qwen" + ] + }, + { + "repository": "cli/cli", + "url": "https://github.com/cli/cli", + "description": "GitHub\u2019s official command line tool", + "default_branch": "trunk", + "archived": false, + "fork": false, + "pushed_at": "2026-09-19T00:29:50Z", + "license_metadata": "MIT", + "topics": [ + "cli", + "git", + "github-api-v4", + "golang" + ] + }, + { + "repository": "flyteorg/flyte", + "url": "https://github.com/flyteorg/flyte", + "description": "Dynamic, resilient AI orchestration. Coordinate data, models, and compute as you build AI workflows.", + "default_branch": "main", + "archived": false, + "fork": false, + "pushed_at": "2026-09-18T23:39:48Z", + "license_metadata": "Apache-2.0", + "topics": [ + "agentic", + "ai-agents", + "ai-development-tools", + "data-analysis", + "data-science", + "declarative", + "fine-tuning", + "flyte", + "golang", + "grpc", + "hacktoberfest", + "kubernetes", + "llm", + "machine-learning", + "mlops", + "orchestration-engine", + "production", + "python", + "scale", + "workflow" + ] + }, + { + "repository": "livekit/agents", + "url": "https://github.com/livekit/agents", + "description": "A framework for building realtime voice AI agents \ud83e\udd16\ud83c\udf99\ufe0f\ud83d\udcf9 ", + "default_branch": "main", + "archived": false, + "fork": false, + "pushed_at": "2026-09-19T00:23:57Z", + "license_metadata": "Apache-2.0", + "topics": [ + "agents", + "ai", + "openai", + "real-time", + "video", + "voice" + ] + }, + { + "repository": "ayghri/i-have-adhd", + "url": "https://github.com/ayghri/i-have-adhd", + "description": "A skill to stop your coding agent from burying the answer. ADHD-friendly output.", + "default_branch": "main", + "archived": false, + "fork": false, + "pushed_at": "2026-09-19T16:44:46Z", + "license_metadata": "MIT", + "topics": [ + "adhd", + "claude-", + "claude-code-plugin", + "claude-skills", + "developer-tools", + "productivity" + ] + }, + { + "repository": "unclebob/swarm-forge", + "url": "https://github.com/unclebob/swarm-forge", + "description": "A simple tool for coordinating several AI agents.", + "default_branch": "main", + "archived": false, + "fork": false, + "pushed_at": "2026-09-07T14:45:22Z", + "license_metadata": null, + "topics": [] + }, + { + "repository": "omacom/omarchy", + "url": "https://github.com/omacom/omarchy", + "description": "Beautiful, Modern & Opinionated Linux", + "default_branch": "quattro", + "archived": false, + "fork": false, + "pushed_at": "2026-09-19T19:53:11Z", + "license_metadata": "MIT", + "topics": [] + }, + { + "repository": "public-apis/public-apis", + "url": "https://github.com/public-apis/public-apis", + "description": "A collective list of free APIs", + "default_branch": "master", + "archived": false, + "fork": false, + "pushed_at": "2026-09-18T19:39:43Z", + "license_metadata": "MIT", + "topics": [ + "api", + "apis", + "dataset", + "development", + "free", + "list", + "lists", + "open-source", + "public", + "public-api", + "public-apis", + "resources", + "software" + ] + }, + { + "repository": "awesome-foss/awesome-sysadmin", + "url": "https://github.com/awesome-foss/awesome-sysadmin", + "description": "A curated list of amazingly awesome open-source sysadmin resources.", + "default_branch": "master", + "archived": false, + "fork": false, + "pushed_at": "2026-09-17T15:16:20Z", + "license_metadata": "NOASSERTION", + "topics": [ + "awesome", + "awesome-list", + "devops", + "list", + "ops", + "self-hosted", + "software", + "sre", + "sysadmin" + ] + }, + { + "repository": "avelino/awesome-go", + "url": "https://github.com/avelino/awesome-go", + "description": "A curated list of awesome Go frameworks, libraries and software", + "default_branch": "main", + "archived": false, + "fork": false, + "pushed_at": "2026-09-20T00:57:41Z", + "license_metadata": "MIT", + "topics": [ + "awesome", + "awesome-list", + "go", + "golang", + "golang-library", + "hacktoberfest" + ] + }, + { + "repository": "awesome-selfhosted/awesome-selfhosted", + "url": "https://github.com/awesome-selfhosted/awesome-selfhosted", + "description": "A list of Free Software network services and web applications which can be hosted on your own servers", + "default_branch": "master", + "archived": false, + "fork": false, + "pushed_at": "2026-09-19T17:28:15Z", + "license_metadata": "NOASSERTION", + "topics": [ + "awesome", + "awesome-list", + "cloud", + "free-software", + "hosting", + "privacy", + "self-hosted", + "selfhosted" + ] + }, + { + "repository": "sdras/awesome-actions", + "url": "https://github.com/sdras/awesome-actions", + "description": "A curated list of awesome actions to use on GitHub", + "default_branch": "main", + "archived": false, + "fork": false, + "pushed_at": "2024-09-01T20:32:39Z", + "license_metadata": "CC0-1.0", + "topics": [ + "actions", + "actions-list", + "awesome", + "awesome-list", + "awesome-lists", + "curated-list", + "github", + "github-actions" + ] + }, + { + "repository": "e2b-dev/awesome-ai-agents", + "url": "https://github.com/e2b-dev/awesome-ai-agents", + "description": "A list of AI autonomous agents", + "default_branch": "main", + "archived": false, + "fork": false, + "pushed_at": "2026-08-21T18:52:45Z", + "license_metadata": "NOASSERTION", + "topics": [ + "agent", + "ai", + "artificial-intelligence", + "autogpt", + "autonomous-agents", + "awesome", + "babyagi", + "copilot", + "gpt", + "gpt-4", + "gpt-engineer", + "openai", + "python" + ] + }, + { + "repository": "wilsonfreitas/awesome-quant", + "url": "https://github.com/wilsonfreitas/awesome-quant", + "description": "A curated list of insanely awesome libraries, packages and resources for Quants (Quantitative Finance)", + "default_branch": "main", + "archived": false, + "fork": false, + "pushed_at": "2026-09-20T01:19:40Z", + "license_metadata": null, + "topics": [ + "algorithmic-trading-engine", + "algorithmic-trading-library", + "algotrading", + "arbitrage-bot", + "awesome", + "awesome-list", + "finance", + "finance-api", + "financial-data", + "financial-instruments", + "google-finance", + "quant", + "quantitative-finance", + "quantitative-trading", + "stock-data", + "technical-analysis", + "trading-algorithms", + "trading-bot", + "trading-strategies", + "yahoo-finance" + ] + }, + { + "repository": "abhisheknaiidu/awesome-github-profile-readme", + "url": "https://github.com/abhisheknaiidu/awesome-github-profile-readme", + "description": "\ud83d\ude0e A curated list of awesome GitHub Profile which updates in real time ", + "default_branch": "master", + "archived": false, + "fork": false, + "pushed_at": "2026-09-11T23:08:19Z", + "license_metadata": "CC0-1.0", + "topics": [ + "awesome", + "awesome-list", + "github", + "github-profile-readme", + "github-readme", + "portfolio", + "profile-readme" + ] + }, + { + "repository": "veggiemonk/awesome-docker", + "url": "https://github.com/veggiemonk/awesome-docker", + "description": ":whale: A curated list of Docker resources and projects", + "default_branch": "master", + "archived": false, + "fork": false, + "pushed_at": "2026-09-12T10:37:31Z", + "license_metadata": "Apache-2.0", + "topics": [ + "awesome", + "awesome-list", + "container", + "docker", + "docker-api", + "docker-container", + "docker-deployment", + "docker-environment", + "docker-image", + "docker-machine", + "docker-monitoring", + "docker-registry", + "docker-security", + "docker-swarm", + "dockerfile", + "list", + "moby", + "tools" + ] + }, + { + "repository": "alebcay/awesome-shell", + "url": "https://github.com/alebcay/awesome-shell", + "description": "A curated list of awesome command-line frameworks, toolkits, guides and gizmos. Inspired by awesome-php.", + "default_branch": "master", + "archived": false, + "fork": false, + "pushed_at": "2025-08-28T14:40:51Z", + "license_metadata": "CC0-1.0", + "topics": [ + "awesome", + "awesome-list", + "bash", + "cli", + "fish", + "list", + "shell", + "zsh" + ] + }, + { + "repository": "github/awesome-copilot", + "url": "https://github.com/github/awesome-copilot", + "description": "Community-contributed instructions, agents, skills, and configurations to help you make the most of GitHub Copilot.", + "default_branch": "main", + "archived": false, + "fork": false, + "pushed_at": "2026-09-19T20:53:58Z", + "license_metadata": "MIT", + "topics": [ + "agent-skills", + "agents", + "ai", + "awesome", + "custom-agents", + "github-copilot", + "hacktoberfest", + "prompt-engineering" + ] + }, + { + "repository": "goabstract/Awesome-Design-Tools", + "url": "https://github.com/goabstract/Awesome-Design-Tools", + "description": "The best design tools and plugins for everything \ud83d\udc49", + "default_branch": "master", + "archived": false, + "fork": false, + "pushed_at": "2024-07-28T19:57:31Z", + "license_metadata": "MIT", + "topics": [ + "animations", + "awesome", + "awesome-list", + "design", + "design-systems", + "design-tools", + "font-awesome", + "ui-design" + ] + }, + { + "repository": "docker/awesome-compose", + "url": "https://github.com/docker/awesome-compose", + "description": "Awesome Docker Compose samples", + "default_branch": "master", + "archived": false, + "fork": false, + "pushed_at": "2026-09-19T01:45:50Z", + "license_metadata": "CC0-1.0", + "topics": [ + "awesome", + "awesome-list", + "docker-compose" + ] + }, + { + "repository": "DovAmir/awesome-design-patterns", + "url": "https://github.com/DovAmir/awesome-design-patterns", + "description": "A curated list of software and architecture related design patterns.", + "default_branch": "master", + "archived": false, + "fork": false, + "pushed_at": "2024-10-25T19:57:00Z", + "license_metadata": null, + "topics": [ + "architecture", + "awesome", + "awesome-list", + "cloud-computing", + "design-patterns", + "gof-patterns", + "lists", + "microservices", + "resources" + ] + }, + { + "repository": "tiimgreen/github-cheat-sheet", + "url": "https://github.com/tiimgreen/github-cheat-sheet", + "description": "A list of cool features of Git and GitHub.", + "default_branch": "master", + "archived": false, + "fork": false, + "pushed_at": "2024-04-15T20:15:56Z", + "license_metadata": "MIT", + "topics": [ + "awesome", + "awesome-list", + "git", + "github", + "list" + ] + }, + { + "repository": "fffaraz/awesome-cpp", + "url": "https://github.com/fffaraz/awesome-cpp", + "description": "A curated list of awesome C++ (or C) frameworks, libraries, resources, and shiny things. Inspired by awesome-... stuff.", + "default_branch": "master", + "archived": false, + "fork": false, + "pushed_at": "2026-09-19T02:12:42Z", + "license_metadata": "MIT", + "topics": [ + "awesome", + "awesome-list", + "c", + "c-plus-plus", + "cpp", + "cpp-library", + "cppcon", + "libraries", + "list", + "lists", + "programming-tutorial", + "resources" + ] + }, + { + "repository": "sindresorhus/awesome-nodejs", + "url": "https://github.com/sindresorhus/awesome-nodejs", + "description": ":zap: Delightful Node.js packages and resources [BECAUSE OF TOO MUCH SPAM AND LOW-QUALITY SUBMISSIONS, SUBMISSIONS ARE PAUSED TEMPORARILY]", + "default_branch": "main", + "archived": false, + "fork": false, + "pushed_at": "2026-09-02T02:11:34Z", + "license_metadata": "CC0-1.0", + "topics": [ + "awesome", + "awesome-list", + "javascript", + "list", + "node", + "nodejs" + ] + }, + { + "repository": "sindresorhus/awesome", + "url": "https://github.com/sindresorhus/awesome", + "description": "\ud83d\ude0e Awesome lists about all kinds of interesting topics [NOTE: Pull requests are temporarily disabled until I have a chance to catch up with the existing ones]", + "default_branch": "main", + "archived": false, + "fork": false, + "pushed_at": "2026-09-02T02:10:47Z", + "license_metadata": "CC0-1.0", + "topics": [ + "awesome", + "awesome-list", + "lists", + "resources", + "unicorns" + ] + }, + { + "repository": "Osmantic/ODS", + "url": "https://github.com/Osmantic/ODS", + "description": "Turn your PC, Mac, or Linux box into an AI server. LLM inference, chat UI, voice, agents, workflows, RAG, and image generation.", + "default_branch": "main", + "archived": false, + "fork": false, + "pushed_at": "2026-09-20T01:42:50Z", + "license_metadata": "Apache-2.0", + "topics": [ + "ai-agents", + "amd", + "comfyui", + "docker", + "llama-cpp", + "llm", + "local-ai", + "n8n", + "nvidia", + "open-webui", + "rag", + "self-hosted", + "speech-to-text", + "strix-halo", + "text-to-speech", + "workflow-automation" + ] + }, + { + "repository": "JetBrains/go-modern-guidelines", + "url": "https://github.com/JetBrains/go-modern-guidelines", + "description": "Help AI coding agents write modern Go", + "default_branch": "main", + "archived": false, + "fork": false, + "pushed_at": "2026-09-10T12:29:04Z", + "license_metadata": "Apache-2.0", + "topics": [ + "ai-agents", + "coding-agent", + "developer-tools", + "go", + "golang", + "guidelines" + ] + }, + { + "repository": "THU-MAIC/OpenMAIC", + "url": "https://github.com/THU-MAIC/OpenMAIC", + "description": "Open Multi-Agent Interactive Classroom \u2014 Get an immersive, multi-agent learning experience in just one click", + "default_branch": "main", + "archived": false, + "fork": false, + "pushed_at": "2026-09-19T20:08:42Z", + "license_metadata": "MIT", + "topics": [] + }, + { + "repository": "bilawalsidhu/gods-eye-view", + "url": "https://github.com/bilawalsidhu/gods-eye-view", + "description": "A spy satellite simulator in your browser, except the data is real. Live open source spatial intelligence on a photorealistic 3D globe.", + "default_branch": "main", + "archived": false, + "fork": false, + "pushed_at": "2026-09-17T00:37:30Z", + "license_metadata": "NOASSERTION", + "topics": [ + "3d-globe", + "cesium", + "flight-tracking", + "geospatial", + "geospatial-intelligence", + "gis", + "osint", + "photogrammetry", + "satellite-tracking", + "spatial-intelligence", + "webgl", + "worldview" + ] + }, + { + "repository": "abi/screenshot-to-code", + "url": "https://github.com/abi/screenshot-to-code", + "description": "Drop in a screenshot and convert it to clean code (HTML/Tailwind/React/Vue)", + "default_branch": "main", + "archived": false, + "fork": false, + "pushed_at": "2026-09-09T17:27:34Z", + "license_metadata": "MIT", + "topics": [] + }, + { + "repository": "bigskysoftware/htmx", + "url": "https://github.com/bigskysoftware/htmx", + "description": " htmx - high power tools for HTML", + "default_branch": "master", + "archived": false, + "fork": false, + "pushed_at": "2026-09-18T00:23:47Z", + "license_metadata": "NOASSERTION", + "topics": [ + "hateoas", + "html", + "htmx", + "hyperscript", + "javascript", + "rest" + ] + }, + { + "repository": "google/googletest", + "url": "https://github.com/google/googletest", + "description": "GoogleTest - Google Testing and Mocking Framework", + "default_branch": "main", + "archived": false, + "fork": false, + "pushed_at": "2026-09-17T17:24:09Z", + "license_metadata": "BSD-3-Clause", + "topics": [] + }, + { + "repository": "weave-os/router", + "url": "https://github.com/weave-os/router", + "description": "Model router for agentic systems. Routes every prompt to the right model in <50ms. Cut costs 40-70% with just an endpoint change.", + "default_branch": "main", + "archived": false, + "fork": false, + "pushed_at": "2026-09-20T00:05:56Z", + "license_metadata": "NOASSERTION", + "topics": [ + "agentic-coding", + "ai-gateway", + "anthropic", + "claude-code", + "codex", + "model-router", + "openai-compatible" + ] + }, + { + "repository": "apache/maka", + "url": "https://github.com/apache/maka", + "description": "Apache Maka (Incubating) is a high-performance agent workspace that keeps a complete record of everything it did.", + "default_branch": "main", + "archived": false, + "fork": false, + "pushed_at": "2026-09-20T01:29:24Z", + "license_metadata": "Apache-2.0", + "topics": [ + "agent-runtime", + "ai", + "ai-agent", + "apache", + "cli", + "desktop", + "electron", + "event-sourcing", + "incubator", + "llm", + "local-first", + "maka", + "tool-use", + "typescript" + ] + }, + { + "repository": "PostHog/posthog", + "url": "https://github.com/PostHog/posthog", + "description": ":hedgehog: PostHog is the leading platform for building self-driving products. Our developer tools \u2013 AI observability, analytics, session replay, flags, experiments, error tracking, logs, and more \u2013 capture all the context agents need to diagnose problems, uncover opportunities, and ship fixes. Steer it all from Slack, web, desktop, or the MCP.", + "default_branch": "master", + "archived": false, + "fork": false, + "pushed_at": "2026-09-20T01:59:26Z", + "license_metadata": "NOASSERTION", + "topics": [ + "ab-testing", + "ai-analytics", + "analytics", + "cdp", + "data-warehouse", + "experiments", + "feature-flags", + "javascript", + "product-analytics", + "python", + "react", + "session-replay", + "surveys", + "typescript", + "web-analytics" + ] + }, + { + "repository": "actions/checkout", + "url": "https://github.com/actions/checkout", + "description": "Action for checking out a repo", + "default_branch": "main", + "archived": false, + "fork": false, + "pushed_at": "2026-09-10T03:02:44Z", + "license_metadata": "MIT", + "topics": [] + }, + { + "repository": "chaitanyagiri/munder-difflin", + "url": "https://github.com/chaitanyagiri/munder-difflin", + "description": "A local multi-agent harness that works with your existing Claude Code, Codex subscriptions, allows you to run an office of agents", + "default_branch": "main", + "archived": false, + "fork": false, + "pushed_at": "2026-09-17T19:12:15Z", + "license_metadata": "MIT", + "topics": [ + "agent-orchestration", + "agents", + "ai-agents", + "autonomous-agents", + "claude-code", + "codex", + "desktop-app", + "electron", + "free", + "gemini-cli", + "harness", + "harness-engineering", + "local-first", + "memory", + "multi-agent", + "opencode", + "orchestration", + "typescript" + ] + }, + { + "repository": "browser-use/browser-use", + "url": "https://github.com/browser-use/browser-use", + "description": "Agents that use the browser.", + "default_branch": "main", + "archived": false, + "fork": false, + "pushed_at": "2026-09-18T22:34:42Z", + "license_metadata": "MIT", + "topics": [ + "ai-agents", + "ai-tools", + "browser-automation", + "browser-use", + "llm", + "playwright", + "python" + ] + }, + { + "repository": "tt-a1i/archify", + "url": "https://github.com/tt-a1i/archify", + "description": "Agent skill for beautiful, verifiable architecture, workflow, sequence, data-flow, and lifecycle diagrams\u2014self-contained HTML with motion and crisp export.", + "default_branch": "main", + "archived": false, + "fork": false, + "pushed_at": "2026-09-19T22:46:53Z", + "license_metadata": "MIT", + "topics": [ + "agent-skills", + "architecture-as-code", + "architecture-diagram", + "claude-skill", + "code-visualization", + "codex", + "coding-agents", + "data-flow-diagram", + "deepseek-harness", + "developer-tools", + "diagram-as-code", + "diagrams", + "diagrams-as-code", + "dsh-plugin", + "mermaid-alternative", + "opencode", + "sequence-diagram", + "software-architecture", + "system-design", + "text-to-diagram" + ] + }, + { + "repository": "VeejaLiu/ScienceFictionCollection", + "url": "https://github.com/VeejaLiu/ScienceFictionCollection", + "description": "\u79d1\u5e7b\u5c0f\u8bf4\u4f5c\u54c1\u96c6\uff0c\u5728\u7ebf\u5730\u5740\uff1ahttps://github-novel-read.vercel.app", + "default_branch": "master", + "archived": false, + "fork": false, + "pushed_at": "2026-08-14T10:24:16Z", + "license_metadata": "MIT", + "topics": [] + }, + { + "repository": "virgiliojr94/book-to-skill", + "url": "https://github.com/virgiliojr94/book-to-skill", + "description": "Turn any technical book PDF into a Claude Code skill \u2014 ready to study, reference, and use while you work.", + "default_branch": "master", + "archived": false, + "fork": false, + "pushed_at": "2026-09-18T13:41:35Z", + "license_metadata": "MIT", + "topics": [ + "agent-skills", + "ai-agents", + "book-to-skill", + "context-engineering", + "document-processing", + "edtech", + "knowledge-base", + "knowledge-management", + "llm", + "pdf-to-markdown", + "rag", + "self-study", + "study-tools" + ] + }, + { + "repository": "bradautomates/claude-video", + "url": "https://github.com/bradautomates/claude-video", + "description": "Give Claude the ability to watch any video. /watch downloads, extracts frames, transcribes, hands it all to Claude.", + "default_branch": "main", + "archived": false, + "fork": false, + "pushed_at": "2026-07-01T01:26:49Z", + "license_metadata": "MIT", + "topics": [] + }, + { + "repository": "donnemartin/system-design-primer", + "url": "https://github.com/donnemartin/system-design-primer", + "description": "Learn how to design large-scale systems. Prep for the system design interview. Includes Anki flashcards.", + "default_branch": "master", + "archived": false, + "fork": false, + "pushed_at": "2026-09-15T01:10:09Z", + "license_metadata": "NOASSERTION", + "topics": [ + "design", + "design-patterns", + "design-system", + "development", + "interview", + "interview-practice", + "interview-questions", + "programming", + "python", + "system", + "web", + "web-application", + "webapp" + ] + }, + { + "repository": "volcengine/OpenViking", + "url": "https://github.com/volcengine/OpenViking", + "description": "Self-evolving Context Database for AI Agents. Unify Agent Memory, Knowledge RAG and Skills.", + "default_branch": "main", + "archived": false, + "fork": false, + "pushed_at": "2026-09-19T20:26:19Z", + "license_metadata": "AGPL-3.0", + "topics": [ + "agent-memory", + "agent-plugins", + "agentic-rag", + "context-database", + "dsh-plugin", + "self-evolving" + ] + }, + { + "repository": "immich-app/immich", + "url": "https://github.com/immich-app/immich", + "description": "High performance self-hosted photo and video management solution.", + "default_branch": "main", + "archived": false, + "fork": false, + "pushed_at": "2026-09-19T17:51:43Z", + "license_metadata": "AGPL-3.0", + "topics": [ + "backup-tool", + "flutter", + "google-photos", + "google-photos-alternative", + "javascript", + "mobile-app", + "nestjs", + "nodejs", + "photo-gallery", + "photos", + "photos-management", + "self-hosted", + "svelte", + "sveltekit", + "typescript", + "videos" + ] + }, + { + "repository": "AlexsJones/llmfit", + "url": "https://github.com/AlexsJones/llmfit", + "description": "Hundreds of models & providers. One command to find what runs on your hardware.", + "default_branch": "main", + "archived": false, + "fork": false, + "pushed_at": "2026-09-19T17:17:14Z", + "license_metadata": "MIT", + "topics": [ + "gguf", + "llm", + "localai", + "mlx", + "skill", + "unsloth" + ] + }, + { + "repository": "akitaonrails/ai-memory", + "url": "https://github.com/akitaonrails/ai-memory", + "description": "Solution for long term memory for agent coding CLIs and to facilitate handoff between different agent vendors", + "default_branch": "main", + "archived": false, + "fork": false, + "pushed_at": "2026-09-20T01:34:23Z", + "license_metadata": "MIT", + "topics": [] + }, + { + "repository": "opengeos/GeoLibre", + "url": "https://github.com/opengeos/GeoLibre", + "description": "A lightweight, cloud-native GIS platform for visualizing, exploring, and analyzing geospatial data. It runs in the web browser, on the desktop, on mobile, and inside Jupyter notebooks.", + "default_branch": "main", + "archived": false, + "fork": false, + "pushed_at": "2026-09-20T00:30:08Z", + "license_metadata": "MIT", + "topics": [ + "cesium", + "data-science", + "duckdb", + "geolibre", + "geospatial", + "gis", + "mapbox", + "maplibre", + "maplibre-gl-js", + "tauri-app" + ] + }, + { + "repository": "paperclipai/paperclip", + "url": "https://github.com/paperclipai/paperclip", + "description": "The open-source app everyone uses to manage agents at work", + "default_branch": "master", + "archived": false, + "fork": false, + "pushed_at": "2026-09-20T02:00:36Z", + "license_metadata": "MIT", + "topics": [] + }, + { + "repository": "cursor/plugins", + "url": "https://github.com/cursor/plugins", + "description": "Cursor plugin specification and official plugins", + "default_branch": "main", + "archived": false, + "fork": false, + "pushed_at": "2026-09-20T00:44:24Z", + "license_metadata": null, + "topics": [] + }, + { + "repository": "cordiverse/cordis", + "url": "https://github.com/cordiverse/cordis", + "description": "Meta-Framework of Spatiotemporal Composability", + "default_branch": "main", + "archived": false, + "fork": false, + "pushed_at": "2026-09-08T18:15:04Z", + "license_metadata": "MIT", + "topics": [ + "effect", + "framework", + "nodejs", + "plugin" + ] + }, + { + "repository": "TencentCloud/TencentDB-Agent-Memory", + "url": "https://github.com/TencentCloud/TencentDB-Agent-Memory", + "description": "TencentDB Agent Memory is a team-level memory hub for AI Agents \u2014 turning conversations, docs, and code into four reusable memory assets (Chat Memory, Skill, LLM-Wiki, Code-Graph) that are governed, shared, and equipped across agents and frameworks.", + "default_branch": "feat/server_team", + "archived": false, + "fork": false, + "pushed_at": "2026-09-18T08:25:50Z", + "license_metadata": "NOASSERTION", + "topics": [ + "agent", + "ai-agent", + "embedding", + "llm", + "local-first", + "long-term-memory", + "memory", + "openclaw-plugin", + "vector-search" + ] + }, + { + "repository": "vitali87/code-graph-rag", + "url": "https://github.com/vitali87/code-graph-rag", + "description": "The ultimate RAG for your monorepo. Query, understand, and edit multi-language codebases with the power of AI and knowledge graphs", + "default_branch": "main", + "archived": false, + "fork": false, + "pushed_at": "2026-09-19T23:40:27Z", + "license_metadata": "MIT", + "topics": [ + "ai", + "ast", + "claude-code", + "code-analysis", + "code-understanding", + "codebase-search", + "developer-tools", + "graph-database", + "knowledge-graph", + "llm", + "mcp", + "mcp-server", + "memgraph", + "monorepo", + "multi-language", + "python", + "rag", + "retrieval-augmented-generation", + "semantic-search", + "tree-sitter" + ] + }, + { + "repository": "NVIDIA-NeMo/Switchyard", + "url": "https://github.com/NVIDIA-NeMo/Switchyard", + "description": "Switchyard lets LLM applications route traffic across models and providers while preserving native OpenAI and Anthropic API compatibility - enabling flexible model selection, benchmarking, and cost/performance optimization.", + "default_branch": "main", + "archived": false, + "fork": false, + "pushed_at": "2026-09-19T15:13:45Z", + "license_metadata": "Apache-2.0", + "topics": [] + }, + { + "repository": "PrimeIntellect-ai/prime-agent", + "url": "https://github.com/PrimeIntellect-ai/prime-agent", + "description": "A self-improving RLM agent for coding workflows and long-running autonomous tasks.", + "default_branch": "main", + "archived": false, + "fork": false, + "pushed_at": "2026-09-19T22:58:16Z", + "license_metadata": "MIT", + "topics": [] + }, + { + "repository": "cathrynlavery/diagram-design", + "url": "https://github.com/cathrynlavery/diagram-design", + "description": "Editorial diagram design for Claude Code, Codex, and Pi. Self-contained HTML + SVG. No shadows. No Mermaid slop.", + "default_branch": "main", + "archived": false, + "fork": false, + "pushed_at": "2026-09-19T17:47:52Z", + "license_metadata": "MIT", + "topics": [ + "agent-skills", + "claude-code", + "codex", + "data-visualization", + "diagrams", + "drawio", + "mermaid", + "svg" + ] + }, + { + "repository": "ankitects/anki", + "url": "https://github.com/ankitects/anki", + "description": "Anki is a smart spaced repetition flashcard program", + "default_branch": "main", + "archived": false, + "fork": false, + "pushed_at": "2026-09-19T01:46:47Z", + "license_metadata": "NOASSERTION", + "topics": [] + }, + { + "repository": "dependabot/dependabot-core", + "url": "https://github.com/dependabot/dependabot-core", + "description": "\ud83e\udd16 Dependabot's core logic for creating update PRs.", + "default_branch": "main", + "archived": false, + "fork": false, + "pushed_at": "2026-09-18T23:13:00Z", + "license_metadata": "MIT", + "topics": [ + "dependencies", + "docker", + "dotnet", + "elixir", + "elm", + "go", + "java", + "javascript", + "php", + "pnpm", + "python", + "ruby", + "rubygems", + "rust", + "terraform" + ] + }, + { + "repository": "usekaneo/kaneo", + "url": "https://github.com/usekaneo/kaneo", + "description": "\ud83c\udfaf All you need. Nothing you don't. Open source project management that works for you, not against you.", + "default_branch": "main", + "archived": false, + "fork": false, + "pushed_at": "2026-09-19T20:53:46Z", + "license_metadata": "MIT", + "topics": [ + "hono", + "issue-management", + "issue-tracker", + "jira-alternative", + "kanban", + "linear-alternative", + "mcp", + "project-management", + "react", + "self-hosted", + "typescript" + ] + }, + { + "repository": "Chachamaru127/claude-code-harness", + "url": "https://github.com/Chachamaru127/claude-code-harness", + "description": "Claude Code Dedicated Development Harness - Achieving High-Quality Development Through an Autonomous Plan\u2192Work\u2192Review Cycle", + "default_branch": "main", + "archived": false, + "fork": false, + "pushed_at": "2026-09-14T01:25:25Z", + "license_metadata": "MIT", + "topics": [] + }, + { + "repository": "VoltAgent/awesome-claude-design", + "url": "https://github.com/VoltAgent/awesome-claude-design", + "description": "Awesome Claude Design: 68 ready-to-use design system inspirations in DESIGN.md format. Drop one in, scaffold a full UI in one shot.", + "default_branch": "main", + "archived": false, + "fork": false, + "pushed_at": "2026-06-20T17:54:18Z", + "license_metadata": "MIT", + "topics": [ + "claude-code", + "claude-design", + "design-md", + "design-system", + "figma" + ] + }, + { + "repository": "SeemSeam/claude_codex_bridge", + "url": "https://github.com/SeemSeam/claude_codex_bridge", + "description": "Visible multi-agent CLI workspace for mixing Codex, Claude, Gemini, Kimi, Qwen, Cursor, Copilot, Pi, OpenCode, and other AI coding agents", + "default_branch": "main", + "archived": false, + "fork": false, + "pushed_at": "2026-09-19T03:44:23Z", + "license_metadata": "NOASSERTION", + "topics": [ + "ai-coding", + "ai-collaboration", + "antigravity", + "claude-code", + "cli", + "codex", + "coding-agent", + "crush", + "cursor-agent", + "gemini", + "github-copilot", + "kimi", + "kiro", + "multi-agent-cli", + "multi-agent-systems", + "opencode", + "pi-coding-agent", + "qwen-code", + "terminal", + "tmux" + ] + }, + { + "repository": "deta/surf", + "url": "https://github.com/deta/surf", + "description": "Personal AI Notebooks. Organize files & webpages and generate notes from them. Open source, local & open data, open model choice (incl. local).", + "default_branch": "main", + "archived": false, + "fork": false, + "pushed_at": "2026-08-24T17:05:49Z", + "license_metadata": "Apache-2.0", + "topics": [ + "claude", + "deepseek", + "gemma", + "knowledge-base", + "knowledge-management", + "llm", + "local", + "local-llm", + "ollama", + "openai", + "productivity", + "rust", + "svelte", + "typescript" + ] + }, + { + "repository": "KnockOutEZ/wigolo", + "url": "https://github.com/KnockOutEZ/wigolo", + "description": "The go-to web for your AI coding agent \u2014 local-first search, fetch, crawl & research over MCP. No API keys, no cloud, $0/query. Public beta.", + "default_branch": "main", + "archived": false, + "fork": false, + "pushed_at": "2026-09-19T09:28:21Z", + "license_metadata": "NOASSERTION", + "topics": [ + "agent", + "ai", + "ai-agent", + "claude", + "cli", + "developer-tools", + "local-first", + "mcp", + "mcp-server", + "metasearch", + "model-context-protocol", + "nodejs", + "privacy", + "rag", + "search", + "search-engine", + "typescript", + "web-crawler", + "web-scraping", + "web-search" + ] + }, + { + "repository": "eugeniughelbur/obsidian-second-brain", + "url": "https://github.com/eugeniughelbur/obsidian-second-brain", + "description": "Persistent memory for Claude Code and 6 other CLI agents, stored as plain markdown in your Obsidian vault. Stop re-explaining your projects, decisions and people every session. 45 commands: hybrid semantic search, self-rewriting notes, key-less web research, and scheduled agents that maintain the vault while you sleep.", + "default_branch": "main", + "archived": false, + "fork": false, + "pushed_at": "2026-09-18T06:48:19Z", + "license_metadata": "MIT", + "topics": [ + "agent-skills", + "ai-agents", + "ai-note-taking", + "ai-second-brain", + "anthropic", + "claude", + "claude-code", + "claude-code-skill", + "claude-memory", + "claude-plugin", + "claude-skill", + "karpathy-llm-wiki", + "knowledge-graph", + "knowledge-management", + "note-taking", + "obsidian", + "obsidian-md", + "personal-knowledge-management", + "pkm", + "second-brain" + ] + }, + { + "repository": "atilaahmettaner/tradingview-mcp", + "url": "https://github.com/atilaahmettaner/tradingview-mcp", + "description": "TradingView MCP server \u2014 real-time market data, technical analysis, screeners & backtesting for Claude, ChatGPT, Cursor & any MCP client. Stocks, crypto, forex & futures across global exchanges. Hosted or self-host.", + "default_branch": "main", + "archived": false, + "fork": false, + "pushed_at": "2026-09-01T15:15:48Z", + "license_metadata": "MIT", + "topics": [ + "claude-desktop", + "cryptocurrency", + "futures", + "market-data", + "mcp-server", + "mcp-tools", + "model-context-protocol", + "openclaw", + "stock-market", + "technical-analysis", + "trading-mcp", + "trading-strategies", + "tradingview" + ] + }, + { + "repository": "0xNyk/council-of-high-intelligence", + "url": "https://github.com/0xNyk/council-of-high-intelligence", + "description": "Structured multi-perspective deliberation for hard decisions. Run full councils, focused triads, or duo debates across Claude Code, Codex, Gemini CLI, and OpenCode.", + "default_branch": "main", + "archived": false, + "fork": false, + "pushed_at": "2026-09-19T10:13:00Z", + "license_metadata": "MIT", + "topics": [ + "agent-skill", + "ai-agents", + "claude-code", + "codex", + "decision-making", + "deliberation", + "gemini-cli", + "llm-routing", + "multi-agent-systems", + "multi-llm", + "open-source", + "opencode", + "prompt-engineering", + "structured-debate" + ] + }, + { + "repository": "open-compass/VLMEvalKit", + "url": "https://github.com/open-compass/VLMEvalKit", + "description": "Open-source evaluation toolkit of large multi-modality models (LMMs), support 220+ LMMs, 80+ benchmarks", + "default_branch": "main", + "archived": false, + "fork": false, + "pushed_at": "2026-09-18T11:51:18Z", + "license_metadata": "Apache-2.0", + "topics": [ + "chatgpt", + "claude", + "clip", + "computer-vision", + "evaluation", + "gemini", + "gpt", + "gpt-4v", + "gpt4", + "large-language-models", + "llava", + "llm", + "multi-modal", + "openai", + "openai-api", + "pytorch", + "qwen", + "vit", + "vqa" + ] + }, + { + "repository": "looplj/axonhub", + "url": "https://github.com/looplj/axonhub", + "description": "\u26a1\ufe0f Open-source AI Gateway \u2014 Use any SDK to call 100+ LLMs. Built-in failover, load balancing, cost control & end-to-end tracing.", + "default_branch": "unstable", + "archived": false, + "fork": false, + "pushed_at": "2026-09-20T00:47:31Z", + "license_metadata": "NOASSERTION", + "topics": [ + "agent", + "agents", + "ai", + "anthropic", + "anthropic-api", + "api-gateway", + "claude", + "claude-code", + "codex", + "cost-management", + "deepseek", + "gemini-api", + "llm", + "openai", + "opencode" + ] + }, + { + "repository": "tradesdontlie/tradingview-mcp", + "url": "https://github.com/tradesdontlie/tradingview-mcp", + "description": "AI-assisted TradingView chart analysis \u2014 connect Claude Code to your TradingView Desktop for personal workflow automation", + "default_branch": "main", + "archived": false, + "fork": false, + "pushed_at": "2026-07-28T17:28:37Z", + "license_metadata": "NOASSERTION", + "topics": [] + }, + { + "repository": "brianpetro/obsidian-smart-connections", + "url": "https://github.com/brianpetro/obsidian-smart-connections", + "description": "Find related notes and excerpts while writing. Your link building copilot displays relevant content in graph + list view. A local embedding model powers semantic search. Zero setup. No API key.", + "default_branch": "main", + "archived": false, + "fork": false, + "pushed_at": "2026-09-16T15:51:49Z", + "license_metadata": "NOASSERTION", + "topics": [ + "chatgpt", + "claude", + "embeddings", + "gemini", + "obsidian", + "obsidian-md", + "obsidian-plugin", + "related-items", + "semantic-search", + "vectors" + ] + }, + { + "repository": "uditgoenka/autoresearch", + "url": "https://github.com/uditgoenka/autoresearch", + "description": "Claude Autoresearch Skill \u2014 Autonomous goal-directed iteration for Claude Code. Inspired by Karpathy's autoresearch. Modify \u2192 Verify \u2192 Keep/Discard \u2192 Repeat forever.", + "default_branch": "master", + "archived": false, + "fork": false, + "pushed_at": "2026-08-12T22:28:18Z", + "license_metadata": "MIT", + "topics": [ + "ai", + "autonomous-agent", + "autoresearch", + "claude", + "claude-code", + "iteration", + "karpathy", + "productivity", + "skill" + ] + }, + { + "repository": "tw93/Waza", + "url": "https://github.com/tw93/Waza", + "description": "\ud83e\udd77 Engineering habits you already know, turned into skills Claude can run.", + "default_branch": "main", + "archived": false, + "fork": false, + "pushed_at": "2026-09-19T14:41:00Z", + "license_metadata": "MIT", + "topics": [ + "claude", + "claude-code", + "des", + "design", + "engineer", + "lear", + "skil", + "skills", + "superpowers" + ] + }, + { + "repository": "open-multi-agent/open-multi-agent", + "url": "https://github.com/open-multi-agent/open-multi-agent", + "description": "Self-hosted TypeScript agent runtime with durable approvals and verifiable run records. Own it, approve it, audit it.", + "default_branch": "main", + "archived": false, + "fork": false, + "pushed_at": "2026-09-18T19:14:04Z", + "license_metadata": "MIT", + "topics": [ + "agent-framework", + "agent-orchestration", + "agentic-ai", + "ai-agents", + "ai-governance", + "anthropic", + "claude", + "claude-code", + "deepseek", + "gemini", + "llm", + "local-llm", + "mcp", + "multi-agent", + "observability", + "ollama", + "openai", + "self-hosted", + "typescript" + ] + }, + { + "repository": "firecrawl/firecrawl-mcp-server", + "url": "https://github.com/firecrawl/firecrawl-mcp-server", + "description": "\ud83d\udd25 Official Firecrawl MCP Server - Adds powerful web scraping and search to Cursor, Claude and any other LLM clients.", + "default_branch": "main", + "archived": false, + "fork": false, + "pushed_at": "2026-09-20T01:59:26Z", + "license_metadata": "MIT", + "topics": [ + "batch-processing", + "claude", + "content-extraction", + "data-collection", + "firecrawl", + "firecrawl-ai", + "javascript-rendering", + "llm-tools", + "mcp", + "mcp-server", + "model-context-protocol", + "search-api", + "web-crawler", + "web-scraping" + ] + }, + { + "repository": "open-compass/opencompass", + "url": "https://github.com/open-compass/opencompass", + "description": "OpenCompass is an LLM evaluation platform, supporting a wide range of models (Llama3, Mistral, InternLM2,GPT-4,LLaMa2, Qwen,GLM, Claude, etc) over 100+ datasets.", + "default_branch": "main", + "archived": false, + "fork": false, + "pushed_at": "2026-09-18T11:52:08Z", + "license_metadata": "Apache-2.0", + "topics": [ + "benchmark", + "chatgpt", + "evaluation", + "large-language-model", + "llama2", + "llama3", + "llm", + "openai" + ] + }, + { + "repository": "di-sukharev/opencommit", + "url": "https://github.com/di-sukharev/opencommit", + "description": "top #1 and most feature rich GPT wrapper for git \u2014 generate commit messages with an LLM in 1 sec \u2014 works with Claude, GPT and every other provider, supports local Ollama models too", + "default_branch": "master", + "archived": false, + "fork": false, + "pushed_at": "2026-09-09T15:32:24Z", + "license_metadata": "MIT", + "topics": [ + "ai", + "ai-commit", + "ai-commits", + "artificial-intelligence", + "chatgpt", + "git", + "gpt", + "productivity" + ] + }, + { + "repository": "NirDiamant/Prompt_Engineering", + "url": "https://github.com/NirDiamant/Prompt_Engineering", + "description": "22 prompt engineering techniques with hands-on Jupyter Notebook tutorials, from fundamental concepts to advanced strategies for leveraging LLMs.", + "default_branch": "main", + "archived": false, + "fork": false, + "pushed_at": "2026-09-19T02:18:01Z", + "license_metadata": "NOASSERTION", + "topics": [ + "ai", + "chain-of-thought", + "chatgpt", + "claude", + "few-shot-learning", + "genai", + "generative-ai", + "gpt", + "in-context-learning", + "langchain", + "llm", + "llms", + "machine-learning", + "openai", + "prompt-engineering", + "prompting", + "python", + "tutorials" + ] + }, + { + "repository": "omnigent-ai/omnigent", + "url": "https://github.com/omnigent-ai/omnigent", + "description": "Omnigent is an open-source AI agent framework and meta-harness: orchestrate Claude Code, Codex, Cursor, Pi, and custom agents \u2014 swap harnesses without rewriting, enforce policies and sandboxing, and collaborate in real time from any device.", + "default_branch": "main", + "archived": false, + "fork": false, + "pushed_at": "2026-09-20T01:15:56Z", + "license_metadata": "Apache-2.0", + "topics": [ + "agent-framework", + "agent-governance", + "agent-orchestration", + "agents", + "ai", + "ai-agent", + "ai-agents", + "claude-code", + "codex", + "coding-agents", + "developer-tools", + "llm", + "ml", + "multi-agent", + "python", + "sandbox" + ] + }, + { + "repository": "google-labs-code/stitch-skills", + "url": "https://github.com/google-labs-code/stitch-skills", + "description": "A library of Agent Skills designed to work with the Stitch MCP server. Each skill follows the Agent Skills open standard, for compatibility with coding agents such as Antigravity, Gemini CLI, Claude Code, Cursor.", + "default_branch": "main", + "archived": false, + "fork": false, + "pushed_at": "2026-08-17T20:20:23Z", + "license_metadata": "Apache-2.0", + "topics": [] + }, + { + "repository": "nexu-io/html-anything", + "url": "https://github.com/nexu-io/html-anything", + "description": "\u2728 The agentic HTML editor \u2014 your local AI agent writes the HTML, you ship it. \ud83d\ude80 75 Skills \u00d7 9 Surfaces (magazine \u00b7 deck \u00b7 poster \u00b7 XHS / tweet \u00b7 prototype \u00b7 data report \u00b7 Hyperframes) \ud83d\udee1\ufe0f Sandboxed preview \u00b7 \ud83d\udce4 1-click to WeChat / X / Zhihu / HTML / PNG \ud83d\udd11 Zero API key \u2014 Claude Code / Cursor / Codex / Gemini / Copilot / OpenCode / Qwen / Aider.", + "default_branch": "main", + "archived": false, + "fork": false, + "pushed_at": "2026-09-15T03:36:24Z", + "license_metadata": "Apache-2.0", + "topics": [ + "agent-skills", + "agentic", + "ai-agents", + "ai-design", + "ai-editor", + "byok", + "claude", + "claude-code", + "claude-skills", + "coding-agents", + "generative-ai", + "html", + "html-editor", + "hyperframes", + "local-first", + "markdown", + "nextjs", + "vibe-coding", + "wechat", + "xiaohongshu" + ] + }, + { + "repository": "cobusgreyling/loop-engineering", + "url": "https://github.com/cobusgreyling/loop-engineering", + "description": "Practical patterns, starters & CLI tools for loop engineering with AI coding agents. Design systems that prompt and orchestrate agents (inspired by Addy Osmani and Boris Cherny). Includes loop-audit, loop-init, loop-cost.", + "default_branch": "main", + "archived": false, + "fork": false, + "pushed_at": "2026-09-19T10:43:04Z", + "license_metadata": "MIT", + "topics": [ + "agentic-ai", + "ai-agents", + "ai-coding", + "anthropic", + "automation", + "claude", + "claude-code", + "codex", + "coding-agents", + "devops-automation", + "devtools", + "github-actions", + "grok", + "llm", + "loop-engineering", + "mcp", + "prompt-engineering" + ] + }, + { + "repository": "mcp-use/mcp-use", + "url": "https://github.com/mcp-use/mcp-use", + "description": "The fullstack MCP framework to develop MCP Apps for ChatGPT / Claude & MCP Servers for AI Agents.", + "default_branch": "main", + "archived": false, + "fork": false, + "pushed_at": "2026-09-18T23:57:03Z", + "license_metadata": "MIT", + "topics": [ + "agent-plugins", + "agentic-framework", + "ai", + "apps-sdk", + "chatgpt", + "claude-code", + "claude-connectors", + "llms", + "mcp", + "mcp-apps", + "mcp-client", + "mcp-gateway", + "mcp-inspector", + "mcp-server", + "mcp-servers", + "mcp-tools", + "mcp-ui", + "model-context-protocol", + "modelcontextprotocol", + "skills" + ] + }, + { + "repository": "1jehuang/jcode", + "url": "https://github.com/1jehuang/jcode", + "description": "The most RAM efficient harness", + "default_branch": "master", + "archived": false, + "fork": false, + "pushed_at": "2026-09-19T06:15:20Z", + "license_metadata": "MIT", + "topics": [ + "ai", + "ai-agent", + "ai-coding-agent", + "claude", + "cli", + "coding-agent", + "llm", + "mcp", + "openai", + "rust", + "terminal", + "tui" + ] + }, + { + "repository": "khoj-ai/khoj", + "url": "https://github.com/khoj-ai/khoj", + "description": "Your AI second brain. Self-hostable. Get answers from the web or your docs. Build custom agents, schedule automations, do deep research. Turn any online or local LLM into your personal, autonomous AI (gpt, claude, gemini, llama, qwen, mistral). Get started - free.", + "default_branch": "master", + "archived": false, + "fork": false, + "pushed_at": "2026-08-02T01:55:40Z", + "license_metadata": "AGPL-3.0", + "topics": [ + "agent", + "ai", + "assistant", + "chat", + "chatgpt", + "emacs", + "image-generation", + "llama3", + "llamacpp", + "llm", + "obsidian", + "obsidian-md", + "offline-llm", + "productivity", + "rag", + "research", + "self-hosted", + "semantic-search", + "stt", + "whatsapp-ai" + ] + }, + { + "repository": "composio-community/awesome-claude-plugins", + "url": "https://github.com/composio-community/awesome-claude-plugins", + "description": "A curated list of Plugins that let you extend Claude Code with custom commands, agents, hooks, and MCP servers through the plugin system.", + "default_branch": "master", + "archived": false, + "fork": false, + "pushed_at": "2026-07-26T09:02:02Z", + "license_metadata": null, + "topics": [ + "anthropic", + "claude-ai", + "claude-code", + "claude-code-plugin", + "claude-code-plugin-marketplace", + "claude-code-plugins", + "claude-code-plugins-marketplace", + "claude-cowork", + "claude-desktop", + "claude-plugins", + "claude-skills", + "plugins" + ] + }, + { + "repository": "Maciek-roboblog/Claude-Code-Usage-Monitor", + "url": "https://github.com/Maciek-roboblog/Claude-Code-Usage-Monitor", + "description": "Real-time Claude Code usage monitor with predictions and warnings", + "default_branch": "main", + "archived": false, + "fork": false, + "pushed_at": "2026-07-05T04:52:21Z", + "license_metadata": "MIT", + "topics": [ + "ai", + "analytics", + "claude", + "claude-code", + "claude-usage", + "limits", + "monitoring", + "terminal", + "usage-tracking" + ] + }, + { + "repository": "zebbern/claude-code-guide", + "url": "https://github.com/zebbern/claude-code-guide", + "description": "Claude Code Guide - Setup, Commands, workflows, agents, skills & tips-n-tricks from beginner to power user!", + "default_branch": "main", + "archived": false, + "fork": false, + "pushed_at": "2026-09-20T01:03:29Z", + "license_metadata": "MIT", + "topics": [ + "ai", + "ai-agent", + "ai-agent-tools", + "anthropic-claude", + "claude", + "claude-ai", + "claude-api", + "claude-code", + "claude-code-communication", + "claude-code-guide", + "claude-code-skills", + "claude-commands", + "claude-desktop", + "claude-mcp", + "claude-sonnet", + "code", + "mcp", + "mcp-agents", + "mcp-tools", + "vscode-extension" + ] + }, + { + "repository": "tradermonty/claude-trading-skills", + "url": "https://github.com/tradermonty/claude-trading-skills", + "description": "Claude Code skills for equity investors and traders \u2014 market analysis, technical charting, economic calendars, screeners, and trading strategy development.", + "default_branch": "main", + "archived": false, + "fork": false, + "pushed_at": "2026-09-20T01:04:25Z", + "license_metadata": "MIT", + "topics": [] + }, + { + "repository": "drona23/claude-token-efficient", + "url": "https://github.com/drona23/claude-token-efficient", + "description": "One CLAUDE.md file. Keeps Claude responses terse. Reduces output verbosity on heavy workflows. Drop-in, no code changes.", + "default_branch": "main", + "archived": false, + "fork": false, + "pushed_at": "2026-06-16T03:10:06Z", + "license_metadata": "MIT", + "topics": [] + }, + { + "repository": "FlorianBruniaux/claude-code-ultimate-guide", + "url": "https://github.com/FlorianBruniaux/claude-code-ultimate-guide", + "description": "The most comprehensive Claude Code guide: agentic workflows, hooks, skills, MCP servers, quizzes, and production-ready templates. 430K+ lines.", + "default_branch": "main", + "archived": false, + "fork": false, + "pushed_at": "2026-09-19T13:58:30Z", + "license_metadata": "CC-BY-SA-4.0", + "topics": [ + "agentic-coding", + "ai-assistant", + "ai-coding", + "ai-pair-programming", + "ai-security", + "anthropic", + "best-practices", + "claude", + "claude-code", + "claude-code-guide", + "claude-code-tutorial", + "cli-tool", + "coding-assistant", + "cursor-alternative", + "developer-tools", + "llm", + "mcp-servers", + "prompt-engineering", + "tutorial", + "vibe-coding" + ] + }, + { + "repository": "awarexone/Agentic-Bug-Hunter", + "url": "https://github.com/awarexone/Agentic-Bug-Hunter", + "description": "AI-powered bug bounty hunting toolkit that works with or without subscription.", + "default_branch": "main", + "archived": false, + "fork": false, + "pushed_at": "2026-09-19T14:41:49Z", + "license_metadata": "MIT", + "topics": [ + "ai-security", + "bug-bounty", + "bugcrowd", + "claude-ai", + "claude-code", + "cti", + "ethical-hacking", + "hackerone", + "hacking", + "hacking-tool", + "penetration-testing", + "recon", + "vulnerability-scanner" + ] + }, + { + "repository": "ykdojo/claude-code-tips", + "url": "https://github.com/ykdojo/claude-code-tips", + "description": "45+ tips for getting the most out of Claude Code, from basics to advanced - includes a custom status line script and Claude Code running itself in a container. Also includes the dx plugin: skills for everyday dev workflows.", + "default_branch": "main", + "archived": false, + "fork": false, + "pushed_at": "2026-09-02T21:51:40Z", + "license_metadata": "NOASSERTION", + "topics": [ + "agentic", + "agentic-ai", + "agentic-coding", + "agentic-workflow", + "ai", + "claude", + "claude-ai", + "claude-code", + "cli", + "developer-tools", + "productivity", + "tips-and-tricks" + ] + }, + { + "repository": "frankbria/ralph-claude-code", + "url": "https://github.com/frankbria/ralph-claude-code", + "description": "Autonomous AI development loop for Claude Code with intelligent exit detection", + "default_branch": "main", + "archived": false, + "fork": false, + "pushed_at": "2026-09-19T02:33:45Z", + "license_metadata": "MIT", + "topics": [ + "ai", + "ai-agent", + "ai-agents", + "ai-development", + "ai-development-tools", + "claude-code", + "claude-code-cli", + "development-tools", + "development-workflow" + ] + }, + { + "repository": "SuperClaude-Org/SuperClaude_Framework", + "url": "https://github.com/SuperClaude-Org/SuperClaude_Framework", + "description": "A configuration framework that enhances Claude Code with specialized commands, cognitive personas, and development methodologies.", + "default_branch": "master", + "archived": false, + "fork": false, + "pushed_at": "2026-09-15T03:10:21Z", + "license_metadata": "MIT", + "topics": [] + }, + { + "repository": "musistudio/claude-code-router", + "url": "https://github.com/musistudio/claude-code-router", + "description": "One local control plane for every AI agent: route across models, fuse new capabilities, orchestrate tools, and stay fully in control.", + "default_branch": "main", + "archived": false, + "fork": false, + "pushed_at": "2026-09-18T02:27:33Z", + "license_metadata": "MIT", + "topics": [] + }, + { + "repository": "anthropics/claude-quickstarts", + "url": "https://github.com/anthropics/claude-quickstarts", + "description": "A collection of projects designed to help developers quickly get started with building deployable applications using the Claude API", + "default_branch": "main", + "archived": false, + "fork": false, + "pushed_at": "2026-09-18T19:15:32Z", + "license_metadata": "MIT", + "topics": [] + }, + { + "repository": "anthropics/claude-plugins-official", + "url": "https://github.com/anthropics/claude-plugins-official", + "description": "Official, Anthropic-managed directory of high quality Claude Code Plugins.", + "default_branch": "main", + "archived": false, + "fork": false, + "pushed_at": "2026-09-18T19:05:53Z", + "license_metadata": "Apache-2.0", + "topics": [ + "claude-code", + "mcp", + "skills" + ] + }, + { + "repository": "catchorg/Catch2", + "url": "https://github.com/catchorg/Catch2", + "description": "A modern, C++-native, test framework for unit-tests, TDD and BDD - using C++14, C++17 and later (C++11 support is in v2.x branch, and C++03 on the Catch1.x branch)", + "default_branch": "devel", + "archived": false, + "fork": false, + "pushed_at": "2026-09-14T12:08:29Z", + "license_metadata": "BSL-1.0", + "topics": [ + "bdd", + "cpp", + "cpp14", + "framework", + "no-dependencies", + "tdd", + "test-framework", + "testing" + ] + }, + { + "repository": "kvcache-ai/ktransformers", + "url": "https://github.com/kvcache-ai/ktransformers", + "description": "A Flexible Framework for Experiencing Heterogeneous LLM Inference/Fine-tune Optimizations", + "default_branch": "main", + "archived": false, + "fork": false, + "pushed_at": "2026-09-19T03:06:55Z", + "license_metadata": "Apache-2.0", + "topics": [] + }, + { + "repository": "earendil-works/pi", + "url": "https://github.com/earendil-works/pi", + "description": "AI agent toolkit: unified LLM API, agent loop, TUI, coding agent CLI", + "default_branch": "main", + "archived": false, + "fork": false, + "pushed_at": "2026-09-19T23:51:04Z", + "license_metadata": "MIT", + "topics": [] + }, + { + "repository": "OtterMind/Chat2DB", + "url": "https://github.com/OtterMind/Chat2DB", + "description": "Chat2DB is a free, cross-platform, local-first database client and SQL workspace for developers, DBAs, analysts, and data teams. Connect to 40+ databases, manage data, edit and run SQL, and use your own AI model to generate, explain, and optimize queries. Available on desktop, web, Docker, and CLI, with MCP support.", + "default_branch": "main", + "archived": false, + "fork": false, + "pushed_at": "2026-09-19T08:45:20Z", + "license_metadata": "NOASSERTION", + "topics": [ + "ai", + "clickhouse", + "database", + "database-client", + "database-gui", + "database-management", + "jdbc", + "llm", + "mcp", + "mongodb", + "mysql", + "oracle", + "postgresql", + "redis", + "sql", + "sql-client", + "sql-editor", + "sql-server", + "sqlite", + "text-to-sql" + ] + }, + { + "repository": "diegosouzapw/OmniRoute", + "url": "https://github.com/diegosouzapw/OmniRoute", + "description": "Never stop coding. Free MIT AI gateway: one endpoint, 352 providers (150+ free), 1200+ models Kimi, Claude, GPT, Gemini, GLM, DeepSeek, MiniMax. Works with Claude Code, Codex, Cursor, OpenCode, Cline & Copilot. Quota-aware auto-fallback, RTK+Caveman compression saves 15-95% tokens, MCP/A2A, Desktop/PWA. Built by 550+ contributors", + "default_branch": "release/v3.8.51", + "archived": false, + "fork": false, + "pushed_at": "2026-09-19T08:30:04Z", + "license_metadata": "MIT", + "topics": [ + "a2a", + "ai-agents", + "ai-gateway", + "anthropic", + "claude", + "claude-code", + "cline", + "codex", + "copilot", + "cursor", + "deepseek", + "free-ai", + "gemini", + "kimi", + "llm-gateway", + "mcp", + "openai", + "openai-proxy", + "qwen", + "token-saver" + ] + }, + { + "repository": "yorukot/superfile", + "url": "https://github.com/yorukot/superfile", + "description": "Pretty fancy and modern terminal file manager", + "default_branch": "main", + "archived": false, + "fork": false, + "pushed_at": "2026-09-18T02:50:00Z", + "license_metadata": "MIT", + "topics": [ + "bubbletea", + "cli", + "file-manager", + "filemanager", + "filesystem", + "golang", + "hacktoberfest", + "linux-app", + "terminal-app", + "terminal-based", + "tui" + ] + }, + { + "repository": "citrolabs/ego-lite", + "url": "https://github.com/citrolabs/ego-lite", + "description": "The fastest browser for AI agents to run browser automation, built for sharing your logged-in browser state with your AI agents, like Codex or Claude Code, without disturbing you. Zero cost, zero config.", + "default_branch": "main", + "archived": false, + "fork": false, + "pushed_at": "2026-09-18T09:27:56Z", + "license_metadata": "MIT", + "topics": [ + "agent-skills", + "ai-agent", + "automation", + "browser", + "browser-automation", + "claude-code", + "codex", + "hermes-agent", + "skills", + "skills-sh" + ] + }, + { + "repository": "likec4/likec4", + "url": "https://github.com/likec4/likec4", + "description": "Visualize, collaborate, and evolve the software architecture with always actual and live diagrams from your code", + "default_branch": "main", + "archived": false, + "fork": false, + "pushed_at": "2026-09-19T22:17:36Z", + "license_metadata": "MIT", + "topics": [ + "architecture", + "architecture-as-code", + "c4", + "diagrams" + ] + }, + { + "repository": "langchain-ai/openwiki", + "url": "https://github.com/langchain-ai/openwiki", + "description": "OpenWiki is a CLI that writes and maintains agent documentation for your codebase.", + "default_branch": "main", + "archived": false, + "fork": false, + "pushed_at": "2026-09-19T08:08:32Z", + "license_metadata": "MIT", + "topics": [] + }, + { + "repository": "nashsu/llm_wiki", + "url": "https://github.com/nashsu/llm_wiki", + "description": "LLM Wiki is a cross-platform desktop application that turns your documents into an organized, interlinked knowledge base \u2014 automatically. Instead of traditional RAG (retrieve-and-answer from scratch every time), the LLM incrementally builds and maintains a persistent wiki from your sources\u3002", + "default_branch": "main", + "archived": false, + "fork": false, + "pushed_at": "2026-08-25T06:42:02Z", + "license_metadata": "NOASSERTION", + "topics": [] + }, + { + "repository": "dgunning/edgartools", + "url": "https://github.com/dgunning/edgartools", + "description": "Read and analyze SEC EDGAR filings in Python. 10-K, 8-K, XBRL financials, Form 3/4/5, 13F, ADV \u2014 clean API, well-typed, MIT-licensed.", + "default_branch": "main", + "archived": false, + "fork": false, + "pushed_at": "2026-09-17T10:29:28Z", + "license_metadata": "MIT", + "topics": [ + "10-k", + "10-q", + "13-f", + "8-k", + "balance-sheet", + "bdc", + "cashflow-statement", + "company", + "cusip", + "edgar", + "filings", + "financials", + "form-4", + "income-statement", + "insider-trading", + "sec", + "tickers", + "xbrl" + ] + }, + { + "repository": "tirth8205/code-review-graph", + "url": "https://github.com/tirth8205/code-review-graph", + "description": "Local-first code intelligence graph for MCP and CLI. Builds a persistent map of your codebase so AI coding tools read only what matters, with benchmarked context reductions on reviews and large-repo workflows.", + "default_branch": "staging", + "archived": false, + "fork": false, + "pushed_at": "2026-09-18T19:41:29Z", + "license_metadata": "MIT", + "topics": [ + "ai-coding", + "claude", + "claude-code", + "code-review", + "graphrag", + "incremental", + "knowledge-graph", + "llm", + "mcp", + "python", + "static-analysis", + "tree-sitter" + ] + }, + { + "repository": "stablyai/orca", + "url": "https://github.com/stablyai/orca", + "description": "Orca is the ADE for working with a fleet of parallel agents. Run any coding agent with your own subscription. Available on desktop, mobile and remote runtime.", + "default_branch": "main", + "archived": false, + "fork": false, + "pushed_at": "2026-09-20T01:29:18Z", + "license_metadata": "MIT", + "topics": [ + "ade", + "agent-ide", + "ai-agents", + "claude-code", + "cli", + "codex", + "cursor-agent", + "devtools", + "ghostty", + "ide", + "mobile-app", + "opencode", + "orchestration", + "parallel-agents", + "pi", + "terminal", + "worktrees", + "yc-backed" + ] + }, + { + "repository": "Zackriya-Solutions/meetily", + "url": "https://github.com/Zackriya-Solutions/meetily", + "description": "Privacy first, AI meeting assistant with 4x faster Parakeet/Whisper live transcription, speaker diarization, and Ollama summarization built on Rust. 100% local processing. no cloud required. Meetily (Meetly Ai - https://meetily.ai) is the #1 Self-hosted, Open-source Ai meeting note taker for macOS & Windows. Understand How to write meeting minutes", + "default_branch": "main", + "archived": false, + "fork": false, + "pushed_at": "2026-09-15T18:32:32Z", + "license_metadata": "MIT", + "topics": [ + "ai", + "ai-meeting-assistant", + "llm", + "local-ai", + "mac", + "meeting-minutes", + "meeting-notes", + "offline-first", + "ollama", + "parakeet", + "privacy-focused", + "privacy-tools", + "rust", + "self-hosted", + "sortformer", + "speech-to-text", + "transcription", + "whisper", + "whisper-cpp", + "windows" + ] + }, + { + "repository": "wonderwhy-er/DesktopCommanderMCP", + "url": "https://github.com/wonderwhy-er/DesktopCommanderMCP", + "description": "This is MCP server for Claude that gives it terminal control, file system search and diff file editing capabilities", + "default_branch": "main", + "archived": false, + "fork": false, + "pushed_at": "2026-09-19T20:30:48Z", + "license_metadata": "MIT", + "topics": [ + "agent", + "ai", + "code-analysis", + "code-generation", + "gemini-cli-extension", + "mcp", + "terminal-ai", + "terminal-automation", + "vibe-coding" + ] + }, + { + "repository": "virattt/ai-hedge-fund", + "url": "https://github.com/virattt/ai-hedge-fund", + "description": "An AI Hedge Fund Team", + "default_branch": "main", + "archived": false, + "fork": false, + "pushed_at": "2026-09-18T14:54:26Z", + "license_metadata": "MIT", + "topics": [] + }, + { + "repository": "ColeMurray/background-agents", + "url": "https://github.com/ColeMurray/background-agents", + "description": "An open-source background agents coding system", + "default_branch": "main", + "archived": false, + "fork": false, + "pushed_at": "2026-09-20T02:00:33Z", + "license_metadata": "MIT", + "topics": [ + "background-agents", + "cloud-agents", + "software-factory" + ] + }, + { + "repository": "Crosstalk-Solutions/project-nomad", + "url": "https://github.com/Crosstalk-Solutions/project-nomad", + "description": "Project NOMAD is an offline-first knowledge and education server. Wikipedia, thousands of books, courses, maps, and optional local AI, all running on hardware you own with no internet required.", + "default_branch": "main", + "archived": false, + "fork": false, + "pushed_at": "2026-09-13T19:33:25Z", + "license_metadata": "Apache-2.0", + "topics": [] + }, + { + "repository": "PrefectHQ/prefect", + "url": "https://github.com/PrefectHQ/prefect", + "description": "Prefect is a workflow orchestration framework for building resilient data pipelines in Python.", + "default_branch": "main", + "archived": false, + "fork": false, + "pushed_at": "2026-09-20T00:24:01Z", + "license_metadata": "Apache-2.0", + "topics": [ + "automation", + "data", + "data-engineering", + "data-ops", + "data-science", + "infrastructure", + "ml-ops", + "observability", + "orchestration", + "pipeline", + "prefect", + "python", + "workflow", + "workflow-engine" + ] + }, + { + "repository": "giancarloerra/SocratiCode", + "url": "https://github.com/giancarloerra/SocratiCode", + "description": "Enterprise-grade (40m+ LOC) codebase intelligence, zero-setup, local & private Plugin/Skill/Extension or MCP: hybrid semantic search, polyglot dependency graphs, symbol-level impact analysis & call-flow, interactive HTML viewer, cross-project & branch-aware search, DB/API/infra knowledge. 61% less tokens, 84% fewer calls, 37x faster. Cloud in beta.", + "default_branch": "main", + "archived": false, + "fork": false, + "pushed_at": "2026-09-17T19:14:39Z", + "license_metadata": "AGPL-3.0", + "topics": [ + "ai", + "ai-assistant", + "ast", + "claude", + "claude-code", + "code-graph", + "codebase-intelligence", + "context-engine", + "docker", + "embeddings", + "gemini", + "gemini-cli-extension", + "mcp", + "openai", + "qdrant", + "semantic", + "semantic-search", + "vector-database", + "vector-embeddings", + "vector-search" + ] + }, + { + "repository": "graykode/abtop", + "url": "https://github.com/graykode/abtop", + "description": "Like htop, but for AI coding agents. Monitor Claude Code & Codex CLI sessions, tokens, context window, rate limits, and ports in real-time.", + "default_branch": "main", + "archived": false, + "fork": false, + "pushed_at": "2026-09-14T04:45:00Z", + "license_metadata": "MIT", + "topics": [ + "ai-agents", + "ai-coding-agent", + "btop", + "claude-code", + "cli", + "codex", + "developer-tools", + "htop", + "monitor", + "ratatui", + "rust", + "terminal", + "tui" + ] + }, + { + "repository": "bytebase/dbhub", + "url": "https://github.com/bytebase/dbhub", + "description": "Token conscious database MCP server for Postgres, MySQL, SQL Server, Oracle, MariaDB, SQLite.", + "default_branch": "main", + "archived": false, + "fork": false, + "pushed_at": "2026-09-19T16:51:07Z", + "license_metadata": "MIT", + "topics": [ + "agents", + "ai", + "anthropic", + "claude", + "claude-ai", + "codex", + "cursor", + "database", + "llm", + "mariadb", + "mcp", + "mcp-server", + "mssql", + "mysql", + "oracle", + "postgres", + "postgresql", + "sql", + "sqlite", + "sqlserver" + ] + }, + { + "repository": "lsdefine/GenericAgent", + "url": "https://github.com/lsdefine/GenericAgent", + "description": "Self-evolving agent: grows skill tree from 3.3K-line seed, achieving full system control with 6x less token consumption", + "default_branch": "main", + "archived": false, + "fork": false, + "pushed_at": "2026-09-19T01:57:31Z", + "license_metadata": "MIT", + "topics": [ + "ai-agent", + "automation", + "autonomous-agent", + "browser-automation", + "claude", + "computer-control", + "desktop-automation", + "gemini", + "lightweight", + "llm-agent", + "memory-system", + "python", + "self-evolving", + "skill-tree", + "task-automation" + ] + }, + { + "repository": "jgravelle/jcodemunch-mcp", + "url": "https://github.com/jgravelle/jcodemunch-mcp", + "description": "Cut AI token costs 95%+ on code exploration. The leading MCP server for precise, symbol-level GitHub code retrieval via tree-sitter AST. Works with Claude Code, Cursor & any MCP client. 313B+ tokens saved.", + "default_branch": "main", + "archived": false, + "fork": false, + "pushed_at": "2026-09-20T01:53:42Z", + "license_metadata": "NOASSERTION", + "topics": [ + "ai-coding", + "ast", + "claude", + "claude-code", + "cline", + "code-intelligence", + "codex", + "context-window", + "copilot", + "cursor", + "developer-tools", + "gemini-cli", + "llm", + "mcp", + "mcp-server", + "model-context-protocol", + "opencode", + "token-optimization", + "tree-sitter", + "windsurf" + ] + }, + { + "repository": "BuilderIO/agent-native", + "url": "https://github.com/BuilderIO/agent-native", + "description": "A framework for building agentic apps", + "default_branch": "main", + "archived": false, + "fork": false, + "pushed_at": "2026-09-20T02:00:38Z", + "license_metadata": null, + "topics": [ + "agent-native", + "agents", + "ai", + "react", + "typescript" + ] + }, + { + "repository": "DeusData/codebase-memory-mcp", + "url": "https://github.com/DeusData/codebase-memory-mcp", + "description": "High-performance code intelligence MCP server. Indexes codebases into a persistent knowledge graph \u2014 average repo in milliseconds. 158 languages, sub-ms queries, 99% fewer tokens. Single static binary, zero dependencies.", + "default_branch": "main", + "archived": false, + "fork": false, + "pushed_at": "2026-09-20T01:50:43Z", + "license_metadata": "MIT", + "topics": [ + "aider", + "ast", + "claude-code", + "code-analysis", + "code-intelligence", + "codex", + "cursor", + "cypher", + "developer-tools", + "gemini-cli", + "graph-visualization", + "kilocode", + "knowledge-graph", + "mcp", + "mcp-server", + "model-context-protocol", + "opencode", + "sqlite", + "tree-sitter", + "windsurf" + ] + }, + { + "repository": "HKUDS/Vibe-Trading", + "url": "https://github.com/HKUDS/Vibe-Trading", + "description": "\"Vibe-Trading: Your Personal Trading Agent\"", + "default_branch": "main", + "archived": false, + "fork": false, + "pushed_at": "2026-09-19T17:10:12Z", + "license_metadata": "MIT", + "topics": [ + "ai-agent", + "algorithmic-trading", + "backtesting", + "fintech", + "llm", + "mcp", + "multi-agent", + "python", + "quantitative-finance", + "trading" + ] + }, + { + "repository": "dbt-labs/dbt", + "url": "https://github.com/dbt-labs/dbt", + "description": "dbt enables data analysts and engineers to transform their data using the same practices that software engineers use to build applications.", + "default_branch": "main", + "archived": false, + "fork": false, + "pushed_at": "2026-09-20T00:59:35Z", + "license_metadata": "Apache-2.0", + "topics": [ + "analytics", + "business-intelligence", + "data-modeling", + "dbt-viewpoint", + "elt", + "pypa", + "slack" + ] + }, + { + "repository": "topoteretes/cognee", + "url": "https://github.com/topoteretes/cognee", + "description": "Cognee is the open-source AI memory platform for agents. Give your AI agents persistent long-term memory across sessions with a self-hosted knowledge graph engine.", + "default_branch": "main", + "archived": false, + "fork": false, + "pushed_at": "2026-09-19T10:54:09Z", + "license_metadata": "Apache-2.0", + "topics": [ + "agent-memory", + "agent-skills", + "ai", + "ai-agents", + "ai-memory", + "cognitive-architecture", + "cognitive-memory", + "context-engineering", + "contributions-welcome", + "good-first-issue", + "good-first-pr", + "graph-database", + "graph-rag", + "help-wanted", + "knowledge", + "knowledge-graph", + "memory-management", + "open-source", + "vector-database" + ] + }, + { + "repository": "ripienaar/free-for-dev", + "url": "https://github.com/ripienaar/free-for-dev", + "description": "A list of SaaS, PaaS and IaaS offerings that have free tiers of interest to devops and infradev", + "default_branch": "master", + "archived": false, + "fork": false, + "pushed_at": "2026-09-18T18:29:20Z", + "license_metadata": null, + "topics": [ + "awesome-list", + "free-for-developers" + ] + }, + { + "repository": "unclecode/crawl4ai", + "url": "https://github.com/unclecode/crawl4ai", + "description": "\ud83d\ude80\ud83e\udd16 Crawl4AI: Open-source LLM Friendly Web Crawler & Scraper. Don't be shy, join here: https://discord.gg/jP8KfhDhyN", + "default_branch": "main", + "archived": false, + "fork": false, + "pushed_at": "2026-09-18T07:00:43Z", + "license_metadata": "Apache-2.0", + "topics": [] + }, + { + "repository": "D4Vinci/Scrapling", + "url": "https://github.com/D4Vinci/Scrapling", + "description": "\ud83d\udd77\ufe0f An adaptive Web Scraping framework that handles everything from a single request to a full-scale crawl! Don't be shy, join here: https://discord.gg/EMgGbDceNQ", + "default_branch": "main", + "archived": false, + "fork": false, + "pushed_at": "2026-09-19T16:15:36Z", + "license_metadata": "BSD-3-Clause", + "topics": [ + "ai", + "ai-scraping", + "automation", + "crawler", + "crawling", + "crawling-python", + "data", + "data-extraction", + "mcp", + "mcp-server", + "playwright", + "python", + "scraping", + "selectors", + "stealth", + "web-scraper", + "web-scraping", + "web-scraping-python", + "webscraping", + "xpath" + ] + }, + { + "repository": "searxng/searxng", + "url": "https://github.com/searxng/searxng", + "description": "SearXNG is a free internet metasearch engine which aggregates results from various search services and databases. Users are neither tracked nor profiled.", + "default_branch": "master", + "archived": false, + "fork": false, + "pushed_at": "2026-09-19T13:05:01Z", + "license_metadata": "AGPL-3.0", + "topics": [ + "bing", + "brave", + "degoogle", + "duckduckgo", + "google", + "metasearch", + "privacy", + "python", + "qwant", + "search", + "search-engine", + "searx", + "searxng", + "startpage", + "yahoo" + ] + }, + { + "repository": "h4ckf0r0day/obscura", + "url": "https://github.com/h4ckf0r0day/obscura", + "description": "The headless browser for AI agents and web scraping", + "default_branch": "main", + "archived": false, + "fork": false, + "pushed_at": "2026-09-19T22:09:25Z", + "license_metadata": "Apache-2.0", + "topics": [ + "antidetect", + "antidetect-browser", + "browser", + "browser-automation", + "cdp", + "headless", + "playwright", + "puppeteer", + "rust" + ] + }, + { + "repository": "trailofbits/skills", + "url": "https://github.com/trailofbits/skills", + "description": "Trail of Bits Claude Code skills for security research, vulnerability detection, and audit workflows", + "default_branch": "main", + "archived": false, + "fork": false, + "pushed_at": "2026-09-16T22:05:12Z", + "license_metadata": "CC-BY-SA-4.0", + "topics": [ + "agent-skills" + ] + }, + { + "repository": "AgriciDaniel/claude-obsidian", + "url": "https://github.com/AgriciDaniel/claude-obsidian", + "description": "Self-organizing AI second brain for Obsidian + Claude Code. Drop any source and Claude reads, links, and files it into one connected knowledge graph of plain Markdown you own. AI note-taking, personal knowledge management (PKM), and an open-source Notion alternative. Based on Karpathy's LLM Wiki pattern.", + "default_branch": "main", + "archived": false, + "fork": false, + "pushed_at": "2026-09-10T17:44:23Z", + "license_metadata": "MIT", + "topics": [ + "agent-skills", + "ai-note-taking", + "ai-second-brain", + "claude-code", + "claude-code-skill", + "claude-memory", + "claude-plugin", + "karpathy-llm-wiki", + "knowledge-graph", + "knowledge-management", + "note-taking", + "notion-alternative", + "obsidian", + "obsidian-ai", + "obsidian-plugin", + "obsidian-second-brain", + "open-source", + "personal-knowledge-management", + "pkm", + "second-brain" + ] + }, + { + "repository": "tursodatabase/turso", + "url": "https://github.com/tursodatabase/turso", + "description": "A SQL database in Rust: SQLite-compatible, now also speaking Postgres (experimental). The LLVM of databases.", + "default_branch": "main", + "archived": false, + "fork": false, + "pushed_at": "2026-09-20T01:53:49Z", + "license_metadata": "MIT", + "topics": [ + "database", + "embedded-database", + "sql", + "sqlite3", + "webassembly" + ] + }, + { + "repository": "agentskills/agentskills", + "url": "https://github.com/agentskills/agentskills", + "description": "Specification and documentation for Agent Skills", + "default_branch": "main", + "archived": false, + "fork": false, + "pushed_at": "2026-08-09T20:36:04Z", + "license_metadata": "Apache-2.0", + "topics": [ + "agent-skills" + ] + }, + { + "repository": "microsoft/SkillOpt", + "url": "https://github.com/microsoft/SkillOpt", + "description": "SkillOpt is a text-space optimizer that trains reusable natural-language skills for frozen LLM agents through trajectory-driven edits, validation-gated updates, and deployable best_skill.md artifacts.", + "default_branch": "main", + "archived": false, + "fork": false, + "pushed_at": "2026-09-05T17:33:48Z", + "license_metadata": "MIT", + "topics": [ + "agent-skills", + "self-evolving-agents" + ] + }, + { + "repository": "mlflow/mlflow", + "url": "https://github.com/mlflow/mlflow", + "description": "The open source AI engineering platform for agents, LLMs, and ML models. MLflow enables teams of all sizes to debug, evaluate, monitor, and optimize production-quality AI applications while controlling costs and managing access to models and data.", + "default_branch": "master", + "archived": false, + "fork": false, + "pushed_at": "2026-09-20T00:55:26Z", + "license_metadata": "Apache-2.0", + "topics": [ + "agentops", + "agents", + "ai", + "ai-governance", + "apache-spark", + "evaluation", + "langchain", + "llm-evaluation", + "llmops", + "machine-learning", + "ml", + "mlflow", + "mlops", + "model-management", + "observability", + "open-source", + "openai", + "prompt-engineering" + ] + }, + { + "repository": "ClickHouse/ClickHouse", + "url": "https://github.com/ClickHouse/ClickHouse", + "description": "ClickHouse\u00ae is a real-time analytics database management system", + "default_branch": "master", + "archived": false, + "fork": false, + "pushed_at": "2026-09-20T01:38:55Z", + "license_metadata": "Apache-2.0", + "topics": [ + "ai", + "analytics", + "big-data", + "clickhouse", + "cloud-native", + "cpp", + "database", + "dbms", + "distributed", + "embedded", + "hacktoberfest", + "lakehouse", + "mpp", + "olap", + "rust", + "self-hosted", + "sql" + ] + }, + { + "repository": "MiroMindAI/MiroThinker", + "url": "https://github.com/MiroMindAI/MiroThinker", + "description": "MiroThinker is a deep research agent optimized for complex research and prediction tasks. Our latest models, MiroThinker-1.7, achieves 74.0 and 75.3 on the BrowseComp and BrowseComp Zh, respectively.", + "default_branch": "main", + "archived": false, + "fork": false, + "pushed_at": "2026-07-06T14:50:39Z", + "license_metadata": "Apache-2.0", + "topics": [ + "agent", + "agent-framework", + "browsecomp", + "deep-research", + "futurex", + "gaia", + "hle", + "research-agent", + "search-agent", + "xbench" + ] + }, + { + "repository": "JuliusBrussee/caveman", + "url": "https://github.com/JuliusBrussee/caveman", + "description": "\ud83e\udea8 why use many token when few token do trick. Viral skill + proxy for coding agents that cuts 65% of tokens by talking like a caveman.", + "default_branch": "main", + "archived": false, + "fork": false, + "pushed_at": "2026-09-20T01:22:02Z", + "license_metadata": "NOASSERTION", + "topics": [ + "ai", + "anthropic", + "caveman", + "claude", + "claude-code", + "llm", + "meme", + "prompt-engineering", + "skill", + "tokens" + ] + }, + { + "repository": "NVIDIA-NeMo/Speech", + "url": "https://github.com/NVIDIA-NeMo/Speech", + "description": "A scalable generative AI framework built for researchers and developers working on Large Language Models, Multimodal, and Speech AI (Automatic Speech Recognition and Text-to-Speech)", + "default_branch": "main", + "archived": false, + "fork": false, + "pushed_at": "2026-09-19T18:56:47Z", + "license_metadata": "Apache-2.0", + "topics": [ + "asr", + "deeplearning", + "generative-ai", + "machine-translation", + "neural-networks", + "speaker-diariazation", + "speaker-recognition", + "speech-synthesis", + "speech-translation", + "tts" + ] + }, + { + "repository": "virattt/dexter", + "url": "https://github.com/virattt/dexter", + "description": "An autonomous agent for deep financial research", + "default_branch": "main", + "archived": false, + "fork": false, + "pushed_at": "2026-08-04T15:20:42Z", + "license_metadata": null, + "topics": [] + }, + { + "repository": "restic/restic", + "url": "https://github.com/restic/restic", + "description": "Fast, secure, efficient backup program", + "default_branch": "master", + "archived": false, + "fork": false, + "pushed_at": "2026-09-01T01:43:58Z", + "license_metadata": "BSD-2-Clause", + "topics": [ + "backup", + "dedupe", + "deduplication", + "go", + "restic", + "secure-by-default" + ] + }, + { + "repository": "pytest-dev/pytest", + "url": "https://github.com/pytest-dev/pytest", + "description": "The pytest framework makes it easy to write small tests, yet scales to support complex functional testing", + "default_branch": "main", + "archived": false, + "fork": false, + "pushed_at": "2026-09-20T00:45:42Z", + "license_metadata": "MIT", + "topics": [ + "hacktoberfest", + "python", + "test", + "testing", + "unit-testing" + ] + }, + { + "repository": "Kong/insomnia", + "url": "https://github.com/Kong/insomnia", + "description": "The open-source, cross-platform API client for GraphQL, REST, WebSockets, SSE and gRPC. With Cloud, Local and Git storage.", + "default_branch": "develop", + "archived": false, + "fork": false, + "pushed_at": "2026-09-18T23:24:25Z", + "license_metadata": "Apache-2.0", + "topics": [ + "api", + "api-client", + "api-design", + "curl", + "electron-app", + "graphql", + "grpc", + "http-client", + "rest-api", + "websockets" + ] + }, + { + "repository": "google-research/timesfm", + "url": "https://github.com/google-research/timesfm", + "description": "TimesFM (Time Series Foundation Model) is a pretrained time-series foundation model developed by Google Research for time-series forecasting.", + "default_branch": "master", + "archived": false, + "fork": false, + "pushed_at": "2026-09-15T20:02:45Z", + "license_metadata": "Apache-2.0", + "topics": [] + }, + { + "repository": "DietrichGebert/ponytail", + "url": "https://github.com/DietrichGebert/ponytail", + "description": "Makes your AI agent think like the laziest senior dev in the room. The best code is the code you never wrote.", + "default_branch": "main", + "archived": false, + "fork": false, + "pushed_at": "2026-09-14T14:34:56Z", + "license_metadata": "MIT", + "topics": [ + "agent-skills", + "ai-agents", + "claude", + "claude-code", + "claude-code-plugin", + "cursor-rules", + "developer-tools", + "llm", + "prompt-engineering", + "yagni" + ] + }, + { + "repository": "SigNoz/signoz", + "url": "https://github.com/SigNoz/signoz", + "description": "SigNoz is an open-source, OpenTelemetry-native observability platform for your team and their AI agents. Get logs, metrics, and traces in one tool with features like APM, distributed tracing, log management, infra monitoring, etc. Combined with SigNoz MCP and a native AI teammate (in SigNoz Cloud) it helps you build more resilient apps.", + "default_branch": "main", + "archived": false, + "fork": false, + "pushed_at": "2026-09-19T15:19:11Z", + "license_metadata": "NOASSERTION", + "topics": [ + "apm", + "application-monitoring", + "distributed-tracing", + "go", + "good-first-issue", + "jaeger", + "log", + "logs", + "metrics", + "monitoring", + "nextjs", + "observability", + "open-source", + "opentelemetry", + "prometheus", + "react", + "reactjs", + "self-hosted", + "tracing", + "typescript" + ] + }, + { + "repository": "Lumiwealth/lumibot", + "url": "https://github.com/Lumiwealth/lumibot", + "description": "AI agents that actually place the trade. 12 brokers, real backtests, stocks options futures forex crypto and prediction markets.", + "default_branch": "dev", + "archived": false, + "fork": false, + "pushed_at": "2026-09-19T17:59:33Z", + "license_metadata": "GPL-3.0", + "topics": [ + "ai-agents", + "algorithmic-trading", + "alpaca", + "backtesting", + "broker", + "crypto-trading", + "finance", + "forex", + "fred", + "futures-trading", + "interactive-brokers", + "llm-agents", + "multi-agent", + "options-trading", + "python", + "quantitative-finance", + "sec-filings", + "stocks", + "technical-analysis", + "trading-bot" + ] + }, + { + "repository": "odysseus-dev/odysseus", + "url": "https://github.com/odysseus-dev/odysseus", + "description": "Self-hosted AI workspace. ", + "default_branch": "dev", + "archived": false, + "fork": false, + "pushed_at": "2026-09-17T17:38:42Z", + "license_metadata": "AGPL-3.0", + "topics": [] + }, + { + "repository": "openai/plugins", + "url": "https://github.com/openai/plugins", + "description": "OpenAI Plugins", + "default_branch": "main", + "archived": false, + "fork": false, + "pushed_at": "2026-09-18T16:58:15Z", + "license_metadata": null, + "topics": [] + }, + { + "repository": "x1xhlol/system-prompts-and-models-of-ai-tools", + "url": "https://github.com/x1xhlol/system-prompts-and-models-of-ai-tools", + "description": "FULL Augment Code, Claude Code, Cluely, CodeBuddy, Comet, Cursor, Devin AI, Junie, Kiro, Leap.new, Lovable, Manus, NotionAI, Orchids.app, Perplexity, Poke, Qoder, Replit, Same.dev, Trae, Traycer AI, VSCode Agent, Warp.dev, Windsurf, Xcode, Z.ai Code, Dia & v0. (And other Open Sourced) System Prompts, Internal Tools & AI Models", + "default_branch": "main", + "archived": false, + "fork": false, + "pushed_at": "2026-08-11T13:01:09Z", + "license_metadata": "GPL-3.0", + "topics": [ + "ai", + "bolt", + "cluely", + "copilot", + "cursor", + "cursorai", + "devin", + "github-copilot", + "lovable", + "open-source", + "perplexity", + "replit", + "system-prompts", + "trae", + "trae-ai", + "trae-ide", + "v0", + "vscode", + "windsurf", + "windsurf-ai" + ] + }, + { + "repository": "microsoft/PowerToys", + "url": "https://github.com/microsoft/PowerToys", + "description": "Microsoft PowerToys is a collection of utilities that supercharge productivity and customization on Windows", + "default_branch": "main", + "archived": false, + "fork": false, + "pushed_at": "2026-09-20T01:11:22Z", + "license_metadata": "MIT", + "topics": [ + "advanced-paste", + "color-picker", + "command-palette", + "desktop", + "fancyzones", + "keyboard-manager", + "microsoft-powertoys", + "powerrename", + "powertoys", + "windows", + "windows-10", + "windows-11" + ] + }, + { + "repository": "LMCache/LMCache", + "url": "https://github.com/LMCache/LMCache", + "description": "LMCache: Supercharge Your LLM with the Fastest KV Cache Layer", + "default_branch": "dev", + "archived": false, + "fork": false, + "pushed_at": "2026-09-20T00:17:07Z", + "license_metadata": "Apache-2.0", + "topics": [ + "amd", + "cuda", + "fast", + "inference", + "kv-cache", + "llm", + "pytorch", + "rocm", + "speed", + "vllm" + ] + }, + { + "repository": "kenn-io/agentsview", + "url": "https://github.com/kenn-io/agentsview", + "description": "Local-first session search, analytics, insights, and token use statistics for coding agents, supporting Claude Code, Codex, and more than 20 other agents. ", + "default_branch": "main", + "archived": false, + "fork": false, + "pushed_at": "2026-09-20T01:34:28Z", + "license_metadata": "MIT", + "topics": [] + }, + { + "repository": "Panniantong/Agent-Reach", + "url": "https://github.com/Panniantong/Agent-Reach", + "description": "Give your AI agent eyes to see the entire internet. Read & search Twitter, Reddit, YouTube, GitHub, Bilibili, XiaoHongShu \u2014 one CLI, zero API fees.", + "default_branch": "main", + "archived": false, + "fork": false, + "pushed_at": "2026-09-15T16:16:24Z", + "license_metadata": "MIT", + "topics": [ + "agent-infrastructure", + "ai-agent", + "ai-search", + "automation", + "bilibili", + "claude-code", + "cli", + "cursor", + "free-api", + "llm-tools", + "mcp", + "python", + "reddit-scraper", + "twitter-scraper", + "web-scraper", + "xiaohongshu", + "youtube-transcript" + ] + }, + { + "repository": "router-for-me/CLIProxyAPI", + "url": "https://github.com/router-for-me/CLIProxyAPI", + "description": "Wrap Antigravity, ChatGPT Codex, Claude Code, Grok Build, Muse Code, Davin as an OpenAI/Gemini/Claude/Codex compatible API service, allowing you to enjoy the free Gemini Series, GPT Series, Grok Series, Claude model through API", + "default_branch": "main", + "archived": false, + "fork": false, + "pushed_at": "2026-09-19T18:58:27Z", + "license_metadata": "MIT", + "topics": [ + "antigravity", + "claude-code", + "cluade", + "codex", + "devin", + "gemini", + "muse", + "openai" + ] + }, + { + "repository": "anthropics/claude-code", + "url": "https://github.com/anthropics/claude-code", + "description": "Claude Code is an agentic coding tool that lives in your terminal, understands your codebase, and helps you code faster by executing routine tasks, explaining complex code, and handling git workflows - all through natural language commands.", + "default_branch": "main", + "archived": false, + "fork": false, + "pushed_at": "2026-09-19T19:16:54Z", + "license_metadata": null, + "topics": [] + }, + { + "repository": "microsoft/markitdown", + "url": "https://github.com/microsoft/markitdown", + "description": "Python tool for converting files and office documents to Markdown.", + "default_branch": "main", + "archived": false, + "fork": false, + "pushed_at": "2026-09-16T17:23:11Z", + "license_metadata": "MIT", + "topics": [ + "autogen", + "autogen-extension", + "langchain", + "markdown", + "microsoft-office", + "openai", + "pdf" + ] + }, + { + "repository": "revfactory/harness", + "url": "https://github.com/revfactory/harness", + "description": "A meta-skill that designs domain-specific agent teams, defines specialized agents, and generates the skills they use.", + "default_branch": "main", + "archived": false, + "fork": false, + "pushed_at": "2026-07-24T23:35:08Z", + "license_metadata": "Apache-2.0", + "topics": [ + "claude-code", + "claude-code-plugin", + "harness", + "harness-engineering" + ] + }, + { + "repository": "lfnovo/open-notebook", + "url": "https://github.com/lfnovo/open-notebook", + "description": "An Open Source implementation of Notebook LM with more flexibility and features", + "default_branch": "main", + "archived": false, + "fork": false, + "pushed_at": "2026-09-18T01:46:29Z", + "license_metadata": "MIT", + "topics": [ + "assistant", + "learning", + "note-taking", + "notebook", + "notes-app", + "self-learning" + ] + }, + { + "repository": "phuryn/pm-skills", + "url": "https://github.com/phuryn/pm-skills", + "description": "PM Skills Marketplace: 100+ agentic skills, commands, and plugins \u2014 from discovery to strategy, execution, launch, and growth.", + "default_branch": "main", + "archived": false, + "fork": false, + "pushed_at": "2026-09-14T21:15:33Z", + "license_metadata": "MIT", + "topics": [ + "agent-skill-repository", + "agent-skills", + "agentic-skills", + "claude-code-marketplace", + "claude-code-plugins", + "claude-cowork-plugin", + "product-management" + ] + }, + { + "repository": "Andyyyy64/whichllm", + "url": "https://github.com/Andyyyy64/whichllm", + "description": "Find the local LLM that actually runs and performs best on your hardware. Ranked by real, recency-aware benchmarks, not parameter count. One command, run it instantly.", + "default_branch": "main", + "archived": false, + "fork": false, + "pushed_at": "2026-09-19T16:22:49Z", + "license_metadata": "MIT", + "topics": [ + "localllm" + ] + }, + { + "repository": "refactoringhq/tolaria", + "url": "https://github.com/refactoringhq/tolaria", + "description": "Desktop app to manage markdown knowledge bases", + "default_branch": "main", + "archived": false, + "fork": false, + "pushed_at": "2026-09-17T13:42:59Z", + "license_metadata": "AGPL-3.0", + "topics": [] + }, + { + "repository": "codejunkie99/agentic-stack", + "url": "https://github.com/codejunkie99/agentic-stack", + "description": "One brain, many harnesses. Portable .agent/ folder (memory + skills + protocols) that plugs into Claude Code, Cursor, Windsurf, OpenCode, OpenClaw, Hermes, or DIY Python \u2014 and keeps its knowledge when you switch.", + "default_branch": "master", + "archived": false, + "fork": false, + "pushed_at": "2026-09-16T19:37:54Z", + "license_metadata": "Apache-2.0", + "topics": [] + }, + { + "repository": "NousResearch/hermes-agent-self-evolution", + "url": "https://github.com/NousResearch/hermes-agent-self-evolution", + "description": "\u2692 Evolutionary self-improvement for Hermes Agent \u2014 optimize skills, prompts, and code using DSPy + GEPA", + "default_branch": "main", + "archived": false, + "fork": false, + "pushed_at": "2026-06-17T11:53:04Z", + "license_metadata": null, + "topics": [] + }, + { + "repository": "0xNyk/awesome-hermes-agent", + "url": "https://github.com/0xNyk/awesome-hermes-agent", + "description": "Independent directory of useful skills, plugins, memory providers, tools, surfaces, and guides for Nous Research's open-source Hermes Agent.", + "default_branch": "main", + "archived": false, + "fork": false, + "pushed_at": "2026-09-19T10:10:28Z", + "license_metadata": "NOASSERTION", + "topics": [ + "agent-skills", + "ai-agents", + "ai-tools", + "awesome", + "awesome-list", + "hermes-agent", + "mcp", + "memory", + "nous-research", + "skills" + ] + }, + { + "repository": "anthropics/claude-agent-sdk-python", + "url": "https://github.com/anthropics/claude-agent-sdk-python", + "description": null, + "default_branch": "main", + "archived": false, + "fork": false, + "pushed_at": "2026-09-19T05:41:13Z", + "license_metadata": "MIT", + "topics": [] + }, + { + "repository": "openai/openai-agents-python", + "url": "https://github.com/openai/openai-agents-python", + "description": "A lightweight, powerful framework for multi-agent workflows", + "default_branch": "main", + "archived": false, + "fork": false, + "pushed_at": "2026-09-18T21:37:55Z", + "license_metadata": "MIT", + "topics": [ + "agents", + "ai", + "framework", + "harness", + "llm", + "openai", + "python" + ] + }, + { + "repository": "walkinglabs/learn-harness-engineering", + "url": "https://github.com/walkinglabs/learn-harness-engineering", + "description": "Harness engineering beginner tutorial, from 0 to 1", + "default_branch": "main", + "archived": false, + "fork": false, + "pushed_at": "2026-08-26T01:08:59Z", + "license_metadata": "MIT", + "topics": [ + "agent", + "agentic", + "agentic-ai", + "ai", + "ai-agent", + "ai-agents", + "dsh", + "dsh-plugin", + "harness", + "harness-engineering", + "harness-framework", + "llm" + ] + }, + { + "repository": "headroomlabs-ai/headroom", + "url": "https://github.com/headroomlabs-ai/headroom", + "description": "Compress tool outputs, logs, files, and RAG chunks before they reach the LLM. 20% fewer tokens for coding agents, 60-95% fewer tokens for JSON, same answers. Library, proxy, MCP server.", + "default_branch": "main", + "archived": false, + "fork": false, + "pushed_at": "2026-09-19T19:07:21Z", + "license_metadata": "Apache-2.0", + "topics": [ + "agent", + "ai", + "anthropic", + "claude-code", + "compression", + "context-engineering", + "context-window", + "cursor", + "fastapi", + "langchain", + "llm", + "mcp", + "openai", + "prompt-engineering", + "proxy", + "python", + "rag", + "token-optimization", + "tokens", + "typescript" + ] + }, + { + "repository": "modelcontextprotocol/modelcontextprotocol", + "url": "https://github.com/modelcontextprotocol/modelcontextprotocol", + "description": "Specification and\u00a0documentation for the Model Context Protocol", + "default_branch": "main", + "archived": false, + "fork": false, + "pushed_at": "2026-09-19T11:01:53Z", + "license_metadata": "NOASSERTION", + "topics": [ + "docs", + "mcp", + "specification", + "standard" + ] + }, + { + "repository": "modelcontextprotocol/servers", + "url": "https://github.com/modelcontextprotocol/servers", + "description": "Model Context Protocol Servers", + "default_branch": "main", + "archived": false, + "fork": false, + "pushed_at": "2026-09-03T01:42:26Z", + "license_metadata": "NOASSERTION", + "topics": [] + }, + { + "repository": "mvanhorn/last30days-skill", + "url": "https://github.com/mvanhorn/last30days-skill", + "description": "AI agent skill that researches any topic across Reddit, X, YouTube, HN, Polymarket, and the web - then synthesizes a grounded summary", + "default_branch": "main", + "archived": false, + "fork": false, + "pushed_at": "2026-09-19T20:00:45Z", + "license_metadata": "MIT", + "topics": [ + "ai-prompts", + "ai-skill", + "bluesky", + "claude", + "claude-code", + "clawhub", + "deep-research", + "hackernews", + "instagram", + "openclaw", + "polymarket", + "recency", + "reddit", + "research", + "social-media", + "tiktok", + "trends", + "twitter", + "web-search", + "youtube" + ] + }, + { + "repository": "aquasecurity/trivy", + "url": "https://github.com/aquasecurity/trivy", + "description": "Find vulnerabilities, misconfigurations, secrets, SBOM in containers, Kubernetes, code repositories, clouds and more", + "default_branch": "main", + "archived": false, + "fork": false, + "pushed_at": "2026-09-18T12:42:01Z", + "license_metadata": "Apache-2.0", + "topics": [ + "containers", + "devsecops", + "docker", + "go", + "golang", + "hacktoberfest", + "iac", + "infrastructure-as-code", + "kubernetes", + "misconfiguration", + "security", + "security-tools", + "vulnerability", + "vulnerability-detection", + "vulnerability-scanners" + ] + }, + { + "repository": "HKUDS/AI-Trader", + "url": "https://github.com/HKUDS/AI-Trader", + "description": "\"AI-Trader: 100% Fully-Automated Agent-Native Trading\" ", + "default_branch": "main", + "archived": false, + "fork": false, + "pushed_at": "2026-06-11T09:26:03Z", + "license_metadata": null, + "topics": [] + }, + { + "repository": "pydantic/pydantic", + "url": "https://github.com/pydantic/pydantic", + "description": "Data validation using Python type hints", + "default_branch": "main", + "archived": false, + "fork": false, + "pushed_at": "2026-09-18T08:12:38Z", + "license_metadata": "MIT", + "topics": [ + "hints", + "json-schema", + "parsing", + "pydantic", + "python", + "python310", + "python311", + "python312", + "python313", + "python39", + "validation" + ] + }, + { + "repository": "VoltAgent/awesome-openclaw-skills", + "url": "https://github.com/VoltAgent/awesome-openclaw-skills", + "description": "The awesome collection of OpenClaw skills. 5,400+ skills filtered and categorized from the official OpenClaw Skills Registry.\ud83e\udd9e", + "default_branch": "main", + "archived": false, + "fork": false, + "pushed_at": "2026-09-14T09:56:58Z", + "license_metadata": "MIT", + "topics": [ + "agent-skills", + "awesome", + "awesome-list", + "awesome-lists", + "clawd", + "clawdbot", + "clawdbot-skill", + "clawdhub", + "moltbot", + "moltbot-skills", + "openclaw", + "openclaw-skills" + ] + }, + { + "repository": "yusufkaraaslan/Skill_Seekers", + "url": "https://github.com/yusufkaraaslan/Skill_Seekers", + "description": "Convert documentation websites, GitHub repositories, and PDFs into Claude AI skills with automatic conflict detection", + "default_branch": "development", + "archived": false, + "fork": false, + "pushed_at": "2026-09-16T20:32:39Z", + "license_metadata": "MIT", + "topics": [ + "ai-tools", + "ast-parser", + "automation", + "claude-ai", + "claude-skills", + "code-analysis", + "conflict-detection", + "documentation", + "documentation-generator", + "github", + "github-scraper", + "mcp", + "mcp-server", + "multi-source", + "ocr", + "pdf", + "python", + "web-scraping" + ] + }, + { + "repository": "VoltAgent/awesome-claude-code-subagents", + "url": "https://github.com/VoltAgent/awesome-claude-code-subagents", + "description": "A collection of 100+ specialized Claude Code subagents covering a wide range of development use cases", + "default_branch": "main", + "archived": false, + "fork": false, + "pushed_at": "2026-09-14T09:59:52Z", + "license_metadata": "MIT", + "topics": [ + "ai-agent-framework", + "ai-agent-tools", + "ai-agents", + "awesome", + "awesome-list", + "claude", + "claude-ai", + "claude-code-subagents", + "claude-subagents", + "subagents" + ] + }, + { + "repository": "oraios/serena", + "url": "https://github.com/oraios/serena", + "description": "A powerful MCP toolkit for coding, providing semantic retrieval and editing capabilities - the IDE for your agent", + "default_branch": "main", + "archived": false, + "fork": false, + "pushed_at": "2026-09-19T10:02:09Z", + "license_metadata": "NOASSERTION", + "topics": [ + "agent", + "ai", + "ai-coding", + "claude", + "claude-code", + "codex", + "ide", + "jetbrains", + "language-server", + "mcp-server", + "programming", + "vibe-coding" + ] + }, + { + "repository": "blader/humanizer", + "url": "https://github.com/blader/humanizer", + "description": "Agent skill that removes signs of AI-generated writing from text", + "default_branch": "main", + "archived": false, + "fork": false, + "pushed_at": "2026-09-06T20:26:10Z", + "license_metadata": "MIT", + "topics": [ + "agent-skills", + "ai-writing", + "claude-code", + "codex", + "cursor", + "prompt-engineering", + "writing-tools" + ] + }, + { + "repository": "K-Dense-AI/scientific-agent-skills", + "url": "https://github.com/K-Dense-AI/scientific-agent-skills", + "description": "Turn any AI agent into an AI Scientist. The #1 Agent Skills library for science, used by 190,000+ scientists worldwide. 165 ready-to-use validated skills plus 100+ scientific databases covering biology, chemistry, medicine, and drug discovery. Compatible with Cursor, Claude Code, Codex, Pi, Antigravity, and the open Agent Skills standard.", + "default_branch": "main", + "archived": false, + "fork": false, + "pushed_at": "2026-09-14T09:27:38Z", + "license_metadata": "MIT", + "topics": [ + "agent-skills", + "ai-scientist", + "bioinformatics", + "chemoinformatics", + "claude", + "claude-skills", + "claudecode", + "clinical-research", + "computational-biology", + "data-analysis", + "drug-discovery", + "genomics", + "materials-science", + "metabolomics", + "proteomics", + "scientific-computing", + "scientific-visualization" + ] + }, + { + "repository": "davila7/claude-code-templates", + "url": "https://github.com/davila7/claude-code-templates", + "description": "CLI tool for configuring and monitoring Claude Code", + "default_branch": "main", + "archived": false, + "fork": false, + "pushed_at": "2026-09-20T01:40:38Z", + "license_metadata": "MIT", + "topics": [ + "anthropic", + "anthropic-claude", + "claude", + "claude-code" + ] + }, + { + "repository": "xingkongliang/skills-manager", + "url": "https://github.com/xingkongliang/skills-manager", + "description": "A lightweight desktop app to manage, sync, and organize AI agent skills across 50+ coding tools \u2014 Claude Code, Codex, Cursor, Copilot, Gemini CLI, and more.", + "default_branch": "main", + "archived": false, + "fork": false, + "pushed_at": "2026-09-17T03:16:02Z", + "license_metadata": "MIT", + "topics": [] + }, + { + "repository": "HKUDS/RAG-Anything", + "url": "https://github.com/HKUDS/RAG-Anything", + "description": "\"RAG-Anything: All-in-One RAG Framework\"", + "default_branch": "main", + "archived": false, + "fork": false, + "pushed_at": "2026-09-15T02:37:26Z", + "license_metadata": "MIT", + "topics": [ + "multi-modal-rag", + "retrieval-augmented-generation" + ] + }, + { + "repository": "mindfold-ai/Trellis", + "url": "https://github.com/mindfold-ai/Trellis", + "description": "The best agent harness.", + "default_branch": "main", + "archived": false, + "fork": false, + "pushed_at": "2026-09-11T18:03:19Z", + "license_metadata": "AGPL-3.0", + "topics": [ + "agentic-coding", + "ai-workflow", + "claudecode", + "codex", + "harness" + ] + }, + { + "repository": "humanlayer/humanlayer", + "url": "https://github.com/humanlayer/humanlayer", + "description": "The best way to get AI coding agents to solve hard problems in complex codebases.", + "default_branch": "main", + "archived": false, + "fork": false, + "pushed_at": "2026-06-19T03:27:53Z", + "license_metadata": "NOASSERTION", + "topics": [ + "agents", + "ai", + "amp", + "claude-code", + "codex", + "human-in-the-loop", + "humanlayer", + "llm", + "llms", + "opencode" + ] + }, + { + "repository": "BigPizzaV3/CodexPlusPlus", + "url": "https://github.com/BigPizzaV3/CodexPlusPlus", + "description": "An enhanced tool for CodexApp, striving to make Codex better to use and more comfortable \u4e00\u4e2aCodexApp\u7684\u589e\u5f3a\u5de5\u5177\uff0c\u52aa\u529b\u8ba9Codex\u53d8\u5f97\u66f4\u597d\u7528\u66f4\u8212\u670d", + "default_branch": "main", + "archived": false, + "fork": false, + "pushed_at": "2026-09-16T18:38:51Z", + "license_metadata": "AGPL-3.0", + "topics": [] + }, + { + "repository": "openai/skills", + "url": "https://github.com/openai/skills", + "description": "Skills Catalog for Codex", + "default_branch": "main", + "archived": false, + "fork": false, + "pushed_at": "2026-09-08T20:35:26Z", + "license_metadata": null, + "topics": [] + }, + { + "repository": "farion1231/cc-switch", + "url": "https://github.com/farion1231/cc-switch", + "description": "A cross-platform desktop All-in-One assistant for Claude Code, Codex, OpenCode, OpenClaw, Grok Build & Hermes Agent. Only official website: ccswitch.io", + "default_branch": "main", + "archived": false, + "fork": false, + "pushed_at": "2026-09-17T18:35:01Z", + "license_metadata": "MIT", + "topics": [ + "ai-tools", + "claude-code", + "codex", + "desktop-app", + "grok", + "grokbuild", + "hermes", + "hermes-agent", + "mcp", + "open-source", + "openclaw", + "openclaw-ui", + "opencode", + "pi", + "provider-management", + "rust", + "skills", + "skills-management", + "tauri", + "wsl-support" + ] + }, + { + "repository": "breferrari/obsidian-mind", + "url": "https://github.com/breferrari/obsidian-mind", + "description": "A self-organizing Obsidian vault that gives AI coding agents persistent memory. Claude Code, Codex CLI, Gemini CLI.", + "default_branch": "main", + "archived": false, + "fork": false, + "pushed_at": "2026-09-02T13:32:14Z", + "license_metadata": "MIT", + "topics": [ + "ai-agents", + "claude-code", + "codex-cli", + "gemini-cli", + "knowledge-management", + "obsidian", + "obsidian-vault", + "persistent-memory", + "second-brain", + "template" + ] + }, + { + "repository": "nyldn/claude-octopus", + "url": "https://github.com/nyldn/claude-octopus", + "description": "Run multiple AI models against the same research, design, or coding task. Surface disagreements before you ship.", + "default_branch": "main", + "archived": false, + "fork": false, + "pushed_at": "2026-09-20T00:44:56Z", + "license_metadata": "MIT", + "topics": [ + "ai-agents", + "ai-orchestration", + "claude-code", + "claude-code-plugin", + "codex", + "copilot", + "developer-tools", + "double-diamond", + "gemini", + "multi-ai", + "multi-llm", + "ollama" + ] + }, + { + "repository": "VoltAgent/awesome-codex-subagents", + "url": "https://github.com/VoltAgent/awesome-codex-subagents", + "description": "A collection of 130+ specialized Codex subagents covering a wide range of development use cases.", + "default_branch": "main", + "archived": false, + "fork": false, + "pushed_at": "2026-09-14T10:04:38Z", + "license_metadata": "MIT", + "topics": [ + "ai-agents", + "awesome-list", + "chatgpt", + "codex", + "codex-skills", + "codex-subagents", + "subagents" + ] + }, + { + "repository": "max-sixty/worktrunk", + "url": "https://github.com/max-sixty/worktrunk", + "description": "Worktrunk is a CLI for Git worktree management, designed for parallel AI agent workflows", + "default_branch": "main", + "archived": false, + "fork": false, + "pushed_at": "2026-09-19T16:30:24Z", + "license_metadata": "NOASSERTION", + "topics": [ + "agents", + "claude-code", + "codex", + "developer-tools", + "git", + "worktrees" + ] + }, + { + "repository": "getagentseal/codeburn", + "url": "https://github.com/getagentseal/codeburn", + "description": "Free, local tool to track AI coding token usage and cost across 37 tools and agents (Claude Code, Cursor, Codex, Gemini and more), by model, project, and task. npx codeburn", + "default_branch": "main", + "archived": false, + "fork": false, + "pushed_at": "2026-09-20T01:43:27Z", + "license_metadata": "MIT", + "topics": [ + "ai-coding", + "claude-code", + "cli", + "codex", + "cost-tracking", + "cursor-ide", + "developer-tools", + "menubar", + "observability", + "terminal-ui", + "token-usage", + "web-dashboard" + ] + }, + { + "repository": "UfoMiao/zcf", + "url": "https://github.com/UfoMiao/zcf", + "description": "Zero-Config Code Flow for Claude code & Codex", + "default_branch": "main", + "archived": false, + "fork": false, + "pushed_at": "2026-08-31T03:02:03Z", + "license_metadata": "MIT", + "topics": [ + "agent", + "ai", + "ai-agent", + "bmad-method", + "ccr", + "claude", + "claude-4", + "claude-ai", + "claude-code", + "cli", + "gpt", + "gpt-5", + "llm", + "llm-code", + "nodejs", + "openai", + "prompt", + "typescript", + "workflow", + "zcf" + ] + }, + { + "repository": "smtg-ai/claude-squad", + "url": "https://github.com/smtg-ai/claude-squad", + "description": "Manage multiple AI terminal agents like Claude Code, Codex, OpenCode, and Amp.", + "default_branch": "main", + "archived": false, + "fork": false, + "pushed_at": "2026-08-20T05:09:53Z", + "license_metadata": "AGPL-3.0", + "topics": [ + "claude-code", + "cli", + "codex", + "opencode", + "vibe-coding" + ] + }, + { + "repository": "Yeachan-Heo/oh-my-codex", + "url": "https://github.com/Yeachan-Heo/oh-my-codex", + "description": "OmX - Oh My codeX: Your codex is not alone. Add hooks, agent teams, HUDs, and so much more.", + "default_branch": "main", + "archived": false, + "fork": false, + "pushed_at": "2026-09-19T17:33:59Z", + "license_metadata": "MIT", + "topics": [] + }, + { + "repository": "kepano/obsidian-skills", + "url": "https://github.com/kepano/obsidian-skills", + "description": "Agent skills for Obsidian. Teach your agent to use Obsidian CLI and open formats including Markdown, Bases, JSON Canvas.", + "default_branch": "main", + "archived": false, + "fork": false, + "pushed_at": "2026-09-15T14:43:57Z", + "license_metadata": "MIT", + "topics": [ + "agents", + "agentskills", + "bases", + "claude", + "clawdbot", + "cli", + "codex", + "defuddle", + "hermes", + "jsoncanvas", + "knap", + "markdown", + "md", + "obsidian", + "openclaw", + "opencode", + "skills" + ] + }, + { + "repository": "can1357/oh-my-pi", + "url": "https://github.com/can1357/oh-my-pi", + "description": "\u2325 Coding agent with the IDE wired in", + "default_branch": "main", + "archived": false, + "fork": false, + "pushed_at": "2026-09-20T01:00:18Z", + "license_metadata": "MIT", + "topics": [ + "ai-agent", + "ai-coding-agent", + "anthropic", + "bun", + "claude", + "cli", + "coding-assistant", + "llm", + "mcp", + "multi-provider", + "openai", + "rust", + "terminal", + "tui", + "typescript" + ] + }, + { + "repository": "nautechsystems/nautilus_trader", + "url": "https://github.com/nautechsystems/nautilus_trader", + "description": "Production-grade Rust-native trading engine with deterministic event-driven architecture", + "default_branch": "develop", + "archived": false, + "fork": false, + "pushed_at": "2026-09-20T00:54:50Z", + "license_metadata": "LGPL-3.0", + "topics": [ + "algorithmic-trading-engine", + "artificial-intelligence", + "crypto-trading", + "equity-trading", + "forex", + "futures-trading", + "machine-learning", + "options-trading", + "python", + "rust", + "sports-betting", + "trading", + "trading-platform" + ] + }, + { + "repository": "shanraisshan/codex-cli-best-practice", + "url": "https://github.com/shanraisshan/codex-cli-best-practice", + "description": "from vibe coding to agentic engineering - practice makes codex perfect", + "default_branch": "main", + "archived": false, + "fork": false, + "pushed_at": "2026-06-04T18:46:09Z", + "license_metadata": "MIT", + "topics": [ + "agentic-ai", + "agentic-coding", + "agentic-engineering", + "agentic-workflow", + "ai", + "ai-agents", + "codex", + "codex-ai", + "codex-cli", + "codex-cli-agents", + "codex-cli-best-practices", + "codex-cli-commands", + "codex-cli-skills", + "codex-hooks", + "context-engineering", + "hooks", + "openai", + "pakistan", + "pakistani-developer", + "vibe-coding" + ] + }, + { + "repository": "modelscope/FunASR", + "url": "https://github.com/modelscope/FunASR", + "description": "Open-source speech recognition toolkit for training, inference, streaming ASR, VAD, punctuation, speaker diarization pipelines, and OpenAI-compatible/MCP serving.", + "default_branch": "main", + "archived": false, + "fork": false, + "pushed_at": "2026-09-19T02:42:10Z", + "license_metadata": "MIT", + "topics": [ + "asr", + "audio", + "chinese", + "emotion-recognition", + "funasr", + "mcp-server", + "multilingual-asr", + "openai-compatible-api", + "paraformer", + "punctuation", + "pytorch", + "real-time-asr", + "speaker-diarization", + "speech-recognition", + "speech-to-text", + "streaming-asr", + "transcription", + "vllm", + "voice-activity-detection", + "whisper-alternative" + ] + }, + { + "repository": "Open-LLM-VTuber/Open-LLM-VTuber", + "url": "https://github.com/Open-LLM-VTuber/Open-LLM-VTuber", + "description": "Talk to any LLM with hands-free voice interaction, voice interruption, and Live2D avatar running locally across platforms", + "default_branch": "main", + "archived": false, + "fork": false, + "pushed_at": "2026-05-15T07:18:04Z", + "license_metadata": "NOASSERTION", + "topics": [ + "ai", + "ai-companion", + "ai-vtuber", + "ai-waifu", + "chatbots", + "live2d", + "live2d-web", + "llm", + "neuro-sama", + "ollama" + ] + }, + { + "repository": "supermemoryai/supermemory", + "url": "https://github.com/supermemoryai/supermemory", + "description": "Memory and context engine + app that is extremely fast, scalable, and can be run fully locally. The Memory API for the AI era.", + "default_branch": "main", + "archived": false, + "fork": false, + "pushed_at": "2026-09-18T22:02:30Z", + "license_metadata": "MIT", + "topics": [ + "agent-memory", + "ai-memory", + "cloudflare-kv", + "cloudflare-pages", + "cloudflare-workers", + "drizzle-orm", + "memory", + "postgres", + "remix", + "tailwindcss", + "typescript", + "vite" + ] + }, + { + "repository": "Orchestra-Research/AI-Research-SKILLs", + "url": "https://github.com/Orchestra-Research/AI-Research-SKILLs", + "description": "Comprehensive open-source library of AI research and engineering skills for any AI model. Package the skills and your claude code/codex/gemini agent will be an AI research agent with full horsepower. Maintained by Orchestra Research.", + "default_branch": "main", + "archived": false, + "fork": false, + "pushed_at": "2026-06-16T01:36:46Z", + "license_metadata": "MIT", + "topics": [ + "ai", + "ai-research", + "claude", + "claude-code", + "claude-skills", + "codex", + "gemini", + "gpt-5", + "grpo", + "huggingface", + "machine-leanring", + "megatron", + "skills", + "vllm" + ] + }, + { + "repository": "assafelovic/gpt-researcher", + "url": "https://github.com/assafelovic/gpt-researcher", + "description": "An autonomous agent that conducts deep research on any data using any LLM providers", + "default_branch": "main", + "archived": false, + "fork": false, + "pushed_at": "2026-08-27T19:34:56Z", + "license_metadata": "Apache-2.0", + "topics": [ + "agent", + "ai", + "automation", + "deepresearch", + "llms", + "mcp", + "mcp-server", + "python", + "research", + "search", + "webscraping" + ] + }, + { + "repository": "bytedance/deer-flow", + "url": "https://github.com/bytedance/deer-flow", + "description": "An open-source long-horizon SuperAgent harness that researches, codes, and creates. With the help of sandboxes, memories, tools, skill, subagents and message gateway, it handles different levels of tasks that could take minutes to hours.", + "default_branch": "main", + "archived": false, + "fork": false, + "pushed_at": "2026-09-19T23:43:07Z", + "license_metadata": "MIT", + "topics": [ + "agent", + "agentic", + "agentic-framework", + "agentic-workflow", + "ai", + "ai-agents", + "deep-research", + "harness", + "langchain", + "langgraph", + "langmanus", + "llm", + "multi-agent", + "nodejs", + "podcast", + "python", + "superagent", + "typescript" + ] + }, + { + "repository": "Piebald-AI/claude-code-system-prompts", + "url": "https://github.com/Piebald-AI/claude-code-system-prompts", + "description": "All parts of Claude Code's system prompt, 27 builtin tool descriptions, sub agent prompts (Plan/Explore/Task), utility prompts (CLAUDE.md, compact, statusline, magic docs, WebFetch, Bash cmd, security review, agent creation). Updated for each Claude Code version.", + "default_branch": "main", + "archived": false, + "fork": false, + "pushed_at": "2026-09-19T14:37:57Z", + "license_metadata": "MIT", + "topics": [ + "claude-code", + "claude-code-system-prompts", + "system-prompts" + ] + }, + { + "repository": "teng-lin/notebooklm-py", + "url": "https://github.com/teng-lin/notebooklm-py", + "description": "Unofficial Python API and agentic skill for Google Gemini Notebook. Full programmatic access to NotebookLM's features\u2014including capabilities the web UI doesn't expose\u2014via Python, CLI, and AI agents like Claude Code, Codex, and OpenClaw.", + "default_branch": "main", + "archived": false, + "fork": false, + "pushed_at": "2026-09-18T22:50:52Z", + "license_metadata": "MIT", + "topics": [ + "agentic-skill", + "claude-skills", + "gemini-notebook", + "gemini-notebook-api", + "gemini-notebook-skill", + "google-notebooklm", + "notebooklm", + "notebooklm-api", + "notebooklm-skill", + "openclaw-skills", + "python", + "python-api", + "sdk", + "skills" + ] + }, + { + "repository": "openai/codex-plugin-cc", + "url": "https://github.com/openai/codex-plugin-cc", + "description": "Use Codex from Claude Code to review code or delegate tasks.", + "default_branch": "main", + "archived": false, + "fork": false, + "pushed_at": "2026-07-08T00:17:31Z", + "license_metadata": "Apache-2.0", + "topics": [] + }, + { + "repository": "VoltAgent/awesome-agent-skills", + "url": "https://github.com/VoltAgent/awesome-agent-skills", + "description": "A curated collection of 1000+ agent skills from official dev teams and the community, compatible with Claude Code, Codex, Gemini CLI, Cursor, and more.", + "default_branch": "main", + "archived": false, + "fork": false, + "pushed_at": "2026-09-15T07:12:04Z", + "license_metadata": "MIT", + "topics": [ + "agent-skills", + "ai-agents", + "awesome", + "awesome-list", + "claude-code", + "claude-code-skills", + "claude-skills", + "codex-skills", + "cursor-skills", + "gemini-skills", + "opencode-skills", + "skills" + ] + }, + { + "repository": "yamadashy/repomix", + "url": "https://github.com/yamadashy/repomix", + "description": "\ud83d\udce6 Repomix is a powerful tool that packs your entire repository into a single, AI-friendly file. Perfect for when you need to feed your codebase to Large Language Models (LLMs) or other AI tools like Claude, ChatGPT, DeepSeek, Perplexity, Gemini, Gemma, Llama, Grok, and more.", + "default_branch": "main", + "archived": false, + "fork": false, + "pushed_at": "2026-09-19T21:09:55Z", + "license_metadata": "MIT", + "topics": [ + "ai", + "anthropic", + "artificial-intelligence", + "chatbot", + "chatgpt", + "claude", + "deepseek", + "developer-tools", + "gemini", + "genai", + "generative-ai", + "gpt", + "javascript", + "language-model", + "llama", + "llm", + "mcp", + "nodejs", + "openai", + "typescript" + ] + }, + { + "repository": "Yeachan-Heo/oh-my-claudecode", + "url": "https://github.com/Yeachan-Heo/oh-my-claudecode", + "description": "Teams-first Multi-agent orchestration for Claude Code", + "default_branch": "main", + "archived": false, + "fork": false, + "pushed_at": "2026-09-18T13:15:08Z", + "license_metadata": "MIT", + "topics": [ + "agentic-coding", + "ai-agents", + "automation", + "claude", + "claude-code", + "multi-agent-systems", + "oh-my-opencode", + "opencode", + "parallel-execution", + "vibe-coding" + ] + }, + { + "repository": "sickn33/agentic-awesome-skills", + "url": "https://github.com/sickn33/agentic-awesome-skills", + "description": "AAS Core is the local, agent-first control plane for complete catalog discovery, agent-owned selection, stack validation, and planning, backed by 2,115+ agentic skills. Includes CLI, local MCP, catalog, plugins, and Workbench.", + "default_branch": "main", + "archived": false, + "fork": false, + "pushed_at": "2026-09-19T18:30:54Z", + "license_metadata": "MIT", + "topics": [ + "agent-skills", + "agentic-skills", + "ai-agent-skills", + "ai-agents", + "ai-coding", + "ai-workflows", + "antigravity", + "antigravity-skills", + "claude-code", + "claude-code-skills", + "codex-cli", + "codex-skills", + "cursor", + "cursor-skills", + "developer-tools", + "gemini-cli", + "gemini-skills", + "kiro", + "mcp", + "skill-library" + ] + }, + { + "repository": "code-yeongyu/oh-my-openagent", + "url": "https://github.com/code-yeongyu/oh-my-openagent", + "description": "OmO: Just type \"mass ulw\" keyword with your prompt. Now you are the master of graph engineering.", + "default_branch": "dev", + "archived": false, + "fork": false, + "pushed_at": "2026-09-19T16:45:00Z", + "license_metadata": "NOASSERTION", + "topics": [ + "ai", + "ai-agents", + "anthropic", + "chatgpt", + "claude", + "claude-skills", + "codex", + "cursor", + "gemini", + "ide", + "openai", + "opencode", + "orchestration", + "tui", + "typescript" + ] + }, + { + "repository": "LearningCircuit/local-deep-research", + "url": "https://github.com/LearningCircuit/local-deep-research", + "description": " ~95% on SimpleQA (e.g. Qwen3.6-27B on a 3090). Supports all local and cloud LLMs (llama.cpp, Ollama, Google, ...). 10+ search engines - arXiv, PubMed, your private documents. Everything Local & Encrypted.", + "default_branch": "main", + "archived": false, + "fork": false, + "pushed_at": "2026-09-19T20:23:37Z", + "license_metadata": "MIT", + "topics": [ + "academia", + "anthropic", + "arxiv", + "brave", + "deep-research", + "encryption", + "home-automation", + "homeserver", + "local", + "local-deep-research", + "local-llm", + "mistral", + "ollama", + "openai", + "pubmed", + "research", + "research-tool", + "retrieval-augmented-generation", + "searxng", + "self-hosted" + ] + }, + { + "repository": "DenisSergeevitch/agents-best-practices", + "url": "https://github.com/DenisSergeevitch/agents-best-practices", + "description": "Provider-neutral Agent Skill for Codex, Claude Code, and agentic harness design.", + "default_branch": "main", + "archived": false, + "fork": false, + "pushed_at": "2026-09-19T17:30:40Z", + "license_metadata": "MIT", + "topics": [ + "agent-skill", + "agent-skills", + "agentic-workflows", + "agents", + "ai-agents", + "anthropic", + "claude", + "claude-code", + "codex", + "codex-skill", + "mcp", + "prompt-engineering" + ] + }, + { + "repository": "EveryInc/compound-engineering-plugin", + "url": "https://github.com/EveryInc/compound-engineering-plugin", + "description": "Official Compound Engineering plugin for Claude Code, Codex, Cursor, and more", + "default_branch": "main", + "archived": false, + "fork": false, + "pushed_at": "2026-09-19T19:23:28Z", + "license_metadata": "MIT", + "topics": [ + "compound", + "engineering" + ] + }, + { + "repository": "anthropics/claude-cookbooks", + "url": "https://github.com/anthropics/claude-cookbooks", + "description": "A collection of notebooks/recipes showcasing some fun and effective ways of using Claude.", + "default_branch": "main", + "archived": false, + "fork": false, + "pushed_at": "2026-09-18T19:07:33Z", + "license_metadata": "MIT", + "topics": [] + }, + { + "repository": "paperless-ngx/paperless-ngx", + "url": "https://github.com/paperless-ngx/paperless-ngx", + "description": "A community-supported supercharged document management system: scan, index and archive all your documents", + "default_branch": "dev", + "archived": false, + "fork": false, + "pushed_at": "2026-09-20T00:59:04Z", + "license_metadata": "GPL-3.0", + "topics": [ + "ai", + "angular", + "archiving", + "django", + "dms", + "document-management", + "document-management-system", + "llm", + "machine-learning", + "ocr", + "optical-character-recognition", + "pdf" + ] + }, + { + "repository": "Fincept-Corporation/FinceptTerminal", + "url": "https://github.com/Fincept-Corporation/FinceptTerminal", + "description": "FinceptTerminal is a modern finance application offering advanced market analytics, investment research, and economic data tools, designed for interactive exploration and data-driven decision-making in a user-friendly environment.", + "default_branch": "main", + "archived": false, + "fork": false, + "pushed_at": "2026-09-19T12:41:43Z", + "license_metadata": "NOASSERTION", + "topics": [ + "ai-agents", + "algorithmic-trading", + "bloomberg-terminal", + "cpp", + "finance", + "financial-markets", + "fintech", + "good-first-issue", + "investment", + "investment-research", + "machine-learning", + "opensource", + "python", + "qt", + "quantitative-finance", + "stock-market", + "trading" + ] + }, + { + "repository": "anthropics/knowledge-work-plugins", + "url": "https://github.com/anthropics/knowledge-work-plugins", + "description": "Open source repository of plugins primarily intended for knowledge workers to use in Claude Cowork", + "default_branch": "main", + "archived": false, + "fork": false, + "pushed_at": "2026-09-19T07:30:50Z", + "license_metadata": "Apache-2.0", + "topics": [] + }, + { + "repository": "Untrivial-ai/agent-orchestrator", + "url": "https://github.com/Untrivial-ai/agent-orchestrator", + "description": "Run and supervise teams of coding agents from planning to merge. Any harness (Claude code, codex, +25 more). Desktop, web, mobile, and cloud agents.", + "default_branch": "main", + "archived": false, + "fork": false, + "pushed_at": "2026-09-20T02:00:20Z", + "license_metadata": "Apache-2.0", + "topics": [ + "agent-fleet", + "agent-ide", + "agent-orchestration", + "agent-swarm", + "claude-code", + "codex-cli", + "git-worktrees", + "multi-agent", + "orchestration", + "orchestrator", + "parallel-agents", + "parallel-coding", + "skills" + ] + }, + { + "repository": "Graphify-Labs/graphify", + "url": "https://github.com/Graphify-Labs/graphify", + "description": "Turn any codebase, with its docs, SQL schemas, configs, and PDFs, into a queryable knowledge graph. A /graphify skill for Claude Code, Cursor, Codex, and Gemini CLI: local deterministic AST parsing, every edge explained, no vector store.", + "default_branch": "v8", + "archived": false, + "fork": false, + "pushed_at": "2026-09-18T21:16:28Z", + "license_metadata": "Apache-2.0", + "topics": [ + "ai-agents", + "antigravity", + "ast", + "claude-code", + "code-analysis", + "code-search", + "codex", + "cursor", + "developer-tools", + "gemini", + "graphrag", + "knowledge-graph", + "leiden", + "llm", + "mcp", + "openclaw", + "rag", + "skills", + "tree-sitter" + ] + }, + { + "repository": "OpenHands/OpenHands", + "url": "https://github.com/OpenHands/OpenHands", + "description": "\ud83d\ude4c OpenHands: AI-Driven Development", + "default_branch": "main", + "archived": false, + "fork": false, + "pushed_at": "2026-09-19T18:02:49Z", + "license_metadata": "MIT", + "topics": [ + "agent", + "artificial-intelligence", + "chatgpt", + "claude-ai", + "cli", + "developer-tools", + "gpt", + "llm", + "openai" + ] + }, + { + "repository": "OthmanAdi/planning-with-files", + "url": "https://github.com/OthmanAdi/planning-with-files", + "description": "Persistent file-based planning for AI coding agents and long-running tasks. Crash-proof markdown plans, session recovery after /clear and compaction, per-turn re-injection against context rot, deterministic completion gate. Manus-style. Install from npm, the Claude Code plugin marketplace, or npx skills. Codex, Cursor, OpenCode, 60+ agents.", + "default_branch": "master", + "archived": false, + "fork": false, + "pushed_at": "2026-09-19T18:57:17Z", + "license_metadata": "MIT", + "topics": [ + "agent-skills", + "autonomous-agents", + "claude", + "claude-code", + "claude-code-skills", + "claude-skills", + "codex", + "coding-agent", + "context-engineering", + "context-rot", + "cursor", + "github-copilot", + "hermes-plugin", + "hermes-skill", + "llm-agents", + "long-running-agents", + "manus", + "multi-agent-systems", + "planning", + "session-recovery" + ] + }, + { + "repository": "Egonex-AI/Understand-Anything", + "url": "https://github.com/Egonex-AI/Understand-Anything", + "description": "Graphs that teach > graphs that impress. Turn any code into an interactive knowledge graph you can explore, search, and ask questions about. Works with Claude Code, Codex, Cursor, Copilot, Gemini CLI, and more.", + "default_branch": "main", + "archived": false, + "fork": false, + "pushed_at": "2026-09-12T05:31:43Z", + "license_metadata": "MIT", + "topics": [ + "antigravity-skills", + "business-knowledge", + "claude-code", + "claude-skills", + "codebase-analysis", + "codex", + "codex-skills", + "developer-tools-ai-agent", + "gemini-cli-skills", + "karpathy-llm-wiki", + "knowledge-base", + "knowledge-graph", + "memory", + "opencode-skills", + "pi-agent", + "understandcode", + "vibe-coding" + ] + }, + { + "repository": "ComposioHQ/composio", + "url": "https://github.com/ComposioHQ/composio", + "description": "Composio powers 1000+ toolkits, tool search, context management, authentication, and a sandboxed workbench to help you build AI agents that turn intent into action.", + "default_branch": "next", + "archived": false, + "fork": false, + "pushed_at": "2026-09-18T20:05:47Z", + "license_metadata": "MIT", + "topics": [ + "agentic-ai", + "agents", + "ai", + "ai-agents", + "aiagents", + "developer-tools", + "function-calling", + "gpt-4", + "javascript", + "js", + "llm", + "llmops", + "mcp", + "python", + "remote-mcp-server", + "sse", + "typescript" + ] + }, + { + "repository": "oven-sh/bun", + "url": "https://github.com/oven-sh/bun", + "description": "Incredibly fast JavaScript runtime, bundler, test runner, and package manager \u2013 all in one", + "default_branch": "main", + "archived": false, + "fork": false, + "pushed_at": "2026-09-20T01:48:19Z", + "license_metadata": "NOASSERTION", + "topics": [ + "bun", + "bundler", + "javascript", + "javascriptcore", + "jsx", + "nodejs", + "npm", + "react", + "rust", + "rust-lang", + "transpiler", + "typescript" + ] + }, + { + "repository": "CloakHQ/CloakBrowser", + "url": "https://github.com/CloakHQ/CloakBrowser", + "description": "Stealth Chromium that passes every bot detection test. Drop-in Playwright replacement with source-level fingerprint patches. 30/30 tests passed.", + "default_branch": "main", + "archived": false, + "fork": false, + "pushed_at": "2026-09-18T17:48:44Z", + "license_metadata": "MIT", + "topics": [ + "ai-agents", + "anti-detect", + "antidetect-browser", + "bot-detection", + "browser-automation", + "captcha-bypass", + "chromium", + "cloudflare", + "cloudflare-bypass", + "fingerprint", + "headless-browser", + "playwright", + "puppeteer", + "python", + "recaptcha", + "selenium", + "stealth-browser", + "undetected", + "web-scraping", + "webscraping" + ] + }, + { + "repository": "tinyhumansai/openhuman", + "url": "https://github.com/tinyhumansai/openhuman", + "description": "OpenHuman is an open source agent harness with local-first memory, agent orchestration, and workflows", + "default_branch": "main", + "archived": false, + "fork": false, + "pushed_at": "2026-09-19T21:43:47Z", + "license_metadata": "GPL-3.0", + "topics": [ + "agent-orchestration", + "ai-agents", + "ai-assistant", + "desktop", + "llm", + "local-first", + "mcp", + "personal-ai", + "privacy", + "rust", + "second-brain", + "tauri" + ] + }, + { + "repository": "Imbad0202/academic-research-skills", + "url": "https://github.com/Imbad0202/academic-research-skills", + "description": "Academic Research Skills for Claude Code: research \u2192 write \u2192 review \u2192 revise \u2192 finalize", + "default_branch": "main", + "archived": false, + "fork": false, + "pushed_at": "2026-09-19T14:01:55Z", + "license_metadata": "NOASSERTION", + "topics": [ + "academic-pipeline", + "academic-writing", + "ai-research", + "claude", + "claude-code", + "literature-review", + "peer-review", + "prompt-engineering" + ] + }, + { + "repository": "colbymchenry/codegraph", + "url": "https://github.com/colbymchenry/codegraph", + "description": "Pre-indexed code knowledge graph, auto syncs on code changes, for Claude Code, Codex, Gemini, Cursor, OpenCode, AntiGravity, Kiro, CoPilot, and Hermes Agent \u2014 fewer tokens, fewer tool calls, 100% local", + "default_branch": "main", + "archived": false, + "fork": false, + "pushed_at": "2026-09-16T18:43:52Z", + "license_metadata": "MIT", + "topics": [] + }, + { + "repository": "HKUDS/CLI-Anything", + "url": "https://github.com/HKUDS/CLI-Anything", + "description": "\"CLI-Anything: Making ALL Software Agent-Native\" -- CLI-Hub: https://clianything.cc/", + "default_branch": "main", + "archived": false, + "fork": false, + "pushed_at": "2026-08-21T07:26:58Z", + "license_metadata": "Apache-2.0", + "topics": [] + }, + { + "repository": "rohitg00/ai-engineering-from-scratch", + "url": "https://github.com/rohitg00/ai-engineering-from-scratch", + "description": "Learn it. Build it. Ship it for others.", + "default_branch": "main", + "archived": false, + "fork": false, + "pushed_at": "2026-09-07T11:42:35Z", + "license_metadata": "MIT", + "topics": [ + "agents", + "ai", + "ai-agents", + "ai-engineering", + "computer-vision", + "course", + "deep-learning", + "from-scratch", + "generative-ai", + "llm", + "machine-learning", + "mcp", + "nlp", + "python", + "reinforcement-learning", + "rust", + "swarm-intelligence", + "transformers", + "tutorial", + "typescript" + ] + }, + { + "repository": "rohitg00/agentmemory", + "url": "https://github.com/rohitg00/agentmemory", + "description": "#1 Persistent memory for AI coding agents based on real-world benchmarks", + "default_branch": "main", + "archived": false, + "fork": false, + "pushed_at": "2026-09-14T06:11:28Z", + "license_metadata": "Apache-2.0", + "topics": [ + "agentmemory", + "agents", + "ai", + "claude", + "claudecode", + "codex", + "copilot", + "cursor", + "genai", + "harness", + "hermes", + "memory", + "openclaw" + ] + }, + { + "repository": "punkpeye/awesome-mcp-servers", + "url": "https://github.com/punkpeye/awesome-mcp-servers", + "description": "A collection of MCP servers.", + "default_branch": "main", + "archived": false, + "fork": false, + "pushed_at": "2026-09-15T23:17:39Z", + "license_metadata": "MIT", + "topics": [ + "ai", + "mcp" + ] + }, + { + "repository": "ComposioHQ/awesome-claude-skills", + "url": "https://github.com/ComposioHQ/awesome-claude-skills", + "description": "A curated list of awesome Claude Skills, resources, and tools for customizing Claude AI workflows", + "default_branch": "master", + "archived": false, + "fork": false, + "pushed_at": "2026-09-18T07:16:11Z", + "license_metadata": null, + "topics": [ + "agent-skills", + "ai-agents", + "antigravity", + "automation", + "claude", + "claude-code", + "codex", + "composio", + "cursor", + "developer-tools", + "gemini-cli", + "mcp", + "openai-codex", + "rube", + "saas", + "skill", + "workflow-automation" + ] + }, + { + "repository": "f/prompts.chat", + "url": "https://github.com/f/prompts.chat", + "description": "f.k.a. Awesome ChatGPT Prompts. Share, discover, and collect prompts from the community. Free and open source \u2014 self-host for your organization with complete privacy.", + "default_branch": "main", + "archived": false, + "fork": false, + "pushed_at": "2026-09-09T10:27:05Z", + "license_metadata": "NOASSERTION", + "topics": [ + "ai", + "artificial-intelligence", + "awesome-list", + "chatgpt", + "chatgpt-prompts", + "claude", + "gemini", + "gpt", + "gpt-4", + "llm", + "machine-learning", + "nextjs", + "open-source", + "openai", + "prompt-engineering", + "prompts", + "prompts-chat", + "typescript" + ] + }, + { + "repository": "addyosmani/agent-skills", + "url": "https://github.com/addyosmani/agent-skills", + "description": "Production-grade engineering skills for AI coding agents.", + "default_branch": "main", + "archived": false, + "fork": false, + "pushed_at": "2026-09-18T03:32:22Z", + "license_metadata": "MIT", + "topics": [ + "agent-skills", + "antigravity", + "claude-code", + "codex", + "cursor", + "skills" + ] + }, + { + "repository": "Open-Dev-Society/OpenStock", + "url": "https://github.com/Open-Dev-Society/OpenStock", + "description": "OpenStock is an open-source alternative to expensive market platforms. Track real-time prices, set personalized alerts, and explore detailed company insights \u2014 built openly, for everyone, forever free.", + "default_branch": "main", + "archived": false, + "fork": false, + "pushed_at": "2026-09-19T13:13:19Z", + "license_metadata": "AGPL-3.0", + "topics": [ + "coderabbit", + "inngest", + "nextjs", + "shadcn-ui", + "stock-market", + "tailwindcss" + ] + }, + { + "repository": "akfamily/akshare", + "url": "https://github.com/akfamily/akshare", + "description": "AKShare is an elegant and simple financial data interface library for Python, built for human beings! \u5f00\u6e90\u8d22\u7ecf\u6570\u636e\u63a5\u53e3\u5e93", + "default_branch": "main", + "archived": false, + "fork": false, + "pushed_at": "2026-09-17T03:40:12Z", + "license_metadata": "MIT", + "topics": [ + "academic", + "akshare", + "asset-pricing", + "bond", + "currency", + "data", + "data-analysis", + "data-science", + "datasets", + "economic-data", + "economics", + "finance", + "finance-api", + "financial-data", + "fundamental", + "futures", + "option", + "quant", + "stock" + ] + }, + { + "repository": "ZhuLinsen/daily_stock_analysis", + "url": "https://github.com/ZhuLinsen/daily_stock_analysis", + "description": "LLM \u9a71\u52a8\u7684\u591a\u5e02\u573a\u80a1\u7968\u667a\u80fd\u5206\u6790\u7cfb\u7edf\uff1a\u591a\u6e90\u884c\u60c5\u3001\u5b9e\u65f6\u65b0\u95fb\u3001\u51b3\u7b56\u770b\u677f\u4e0e\u81ea\u52a8\u63a8\u9001\uff0c\u652f\u6301\u96f6\u6210\u672c\u5b9a\u65f6\u8fd0\u884c\u3002 LLM-powered multi-market stock analysis system with multi-source market data, real-time news, decision dashboard, automated notifications, and cost-free scheduled runs.", + "default_branch": "main", + "archived": false, + "fork": false, + "pushed_at": "2026-09-19T15:09:08Z", + "license_metadata": "MIT", + "topics": [ + "a-stock", + "ai-agent", + "aigc", + "llm", + "quant", + "quantitative-finance", + "quantitative-trading" + ] + }, + { + "repository": "wangzhe3224/awesome-systematic-trading", + "url": "https://github.com/wangzhe3224/awesome-systematic-trading", + "description": "A curated list of insanely awesome libraries, packages and resources for systematic trading. Crypto, Stock, Futures, Options, CFDs, FX, and more | \u91cf\u5316\u4ea4\u6613 | \u91cf\u5316\u6295\u8d44", + "default_branch": "master", + "archived": false, + "fork": false, + "pushed_at": "2026-08-30T20:58:30Z", + "license_metadata": "MIT", + "topics": [ + "algorithmic-trading", + "alpha", + "awesome-list", + "backtesting", + "bitcoin", + "cryptocurrencies", + "cryptocurrency", + "finance", + "finances", + "golang", + "python", + "quant", + "quantitative-trading", + "rust", + "systematic-trading", + "systematic-trading-strategies", + "trading", + "trading-algorithms", + "trading-bot", + "trading-strategies" + ] + }, + { + "repository": "rtk-ai/rtk", + "url": "https://github.com/rtk-ai/rtk", + "description": "CLI proxy that reduces LLM token consumption by 60-90% on common dev commands. Single Rust binary, zero dependencies", + "default_branch": "develop", + "archived": false, + "fork": false, + "pushed_at": "2026-09-19T21:14:12Z", + "license_metadata": "Apache-2.0", + "topics": [ + "agentic-coding", + "ai-coding", + "anthropic", + "claude-code", + "cli", + "command-line-tool", + "cost-reduction", + "developer-tools", + "llm", + "open-source", + "productivity", + "rust", + "token-optimization" + ] + }, + { + "repository": "composio-community/awesome-codex-skills", + "url": "https://github.com/composio-community/awesome-codex-skills", + "description": "A curated list of practical Codex skills for automating workflows across the Codex CLI and API.", + "default_branch": "master", + "archived": false, + "fork": false, + "pushed_at": "2026-07-26T01:14:54Z", + "license_metadata": null, + "topics": [ + "awesome", + "awesome-lists", + "awesome-resources", + "codex", + "codex-cli", + "codex-skills", + "coding-agent-skills", + "coding-agents", + "gpt-5-1-codex", + "gpt-5-codex", + "llm", + "skills" + ] + }, + { + "repository": "warpdotdev/warp", + "url": "https://github.com/warpdotdev/warp", + "description": "Warp is an agentic development environment, born out of the terminal.", + "default_branch": "master", + "archived": false, + "fork": false, + "pushed_at": "2026-09-20T01:42:59Z", + "license_metadata": "AGPL-3.0", + "topics": [ + "bash", + "linux", + "macos", + "rust", + "shell", + "terminal", + "wasm", + "zsh" + ] + }, + { + "repository": "sirmalloc/ccstatusline", + "url": "https://github.com/sirmalloc/ccstatusline", + "description": "\ud83d\ude80 Beautiful highly customizable statusline for Claude Code CLI with powerline support, themes, and more.", + "default_branch": "main", + "archived": false, + "fork": false, + "pushed_at": "2026-09-18T17:45:35Z", + "license_metadata": "MIT", + "topics": [ + "ai-tools", + "claude-code", + "cli", + "developer-tools", + "git", + "powerline", + "statusbar", + "statusline", + "terminal", + "themeing" + ] + }, + { + "repository": "vercel-labs/agent-skills", + "url": "https://github.com/vercel-labs/agent-skills", + "description": "Vercel's official collection of agent skills", + "default_branch": "main", + "archived": false, + "fork": false, + "pushed_at": "2026-08-28T13:36:31Z", + "license_metadata": null, + "topics": [] + }, + { + "repository": "multica-ai/multica", + "url": "https://github.com/multica-ai/multica", + "description": "Make humans and AI agents work as one team \u2014 open-source and self-hostable.", + "default_branch": "main", + "archived": false, + "fork": false, + "pushed_at": "2026-09-19T10:18:03Z", + "license_metadata": "NOASSERTION", + "topics": [] + }, + { + "repository": "rowboatlabs/rowboat", + "url": "https://github.com/rowboatlabs/rowboat", + "description": "AI coworker with memory and collaboration", + "default_branch": "main", + "archived": false, + "fork": false, + "pushed_at": "2026-09-18T18:16:08Z", + "license_metadata": "Apache-2.0", + "topics": [ + "agents", + "agents-sdk", + "ai", + "ai-agents", + "ai-agents-automation", + "chatgpt", + "claude-code", + "claude-cowork", + "generative-ai", + "llm", + "multiagent", + "opeani", + "open-source", + "orchestration", + "productivity" + ] + }, + { + "repository": "mksglu/context-mode", + "url": "https://github.com/mksglu/context-mode", + "description": "Context window optimization for AI coding agents. Sandboxes tool output (98% reduction), persists session memory, and enforces routing across 17 platforms via MCP + hooks.", + "default_branch": "main", + "archived": false, + "fork": false, + "pushed_at": "2026-09-19T12:05:41Z", + "license_metadata": "NOASSERTION", + "topics": [ + "antigravity", + "claude", + "claude-code", + "claude-code-hooks", + "claude-code-plugins", + "claude-code-skill", + "codex", + "codex-cli", + "context-mode", + "copilot", + "cursor-plugin", + "kiro", + "mcp", + "mcp-server", + "mcp-tools", + "openclaw", + "opencode", + "pi-agent", + "skills", + "zed-extension" + ] + }, + { + "repository": "zilliztech/claude-context", + "url": "https://github.com/zilliztech/claude-context", + "description": "Code search MCP for Claude Code. Make entire codebase the context for any coding agent.", + "default_branch": "master", + "archived": false, + "fork": false, + "pushed_at": "2026-07-14T12:00:38Z", + "license_metadata": "MIT", + "topics": [ + "agent", + "agentic-rag", + "ai-coding", + "claude-code", + "code-generation", + "code-search", + "cursor", + "embedding", + "gemini-cli", + "mcp", + "merkle-tree", + "nodejs", + "openai", + "rag", + "semantic-search", + "typescript", + "vector-database", + "vibe-coding", + "voyage-ai", + "vscode-extension" + ] + }, + { + "repository": "HKUDS/DeepTutor", + "url": "https://github.com/HKUDS/DeepTutor", + "description": "DeepTutor: Lifelong Personalized Tutoring. https://deeptutor.info/.", + "default_branch": "main", + "archived": false, + "fork": false, + "pushed_at": "2026-09-19T08:40:16Z", + "license_metadata": "Apache-2.0", + "topics": [ + "ai-agents", + "ai-tutor", + "clawdbot", + "cli-tool", + "deepresearch", + "interactive-learning", + "large-language-models", + "multi-agent-systems", + "rag" + ] + }, + { + "repository": "AsyncFuncAI/deepwiki-open", + "url": "https://github.com/AsyncFuncAI/deepwiki-open", + "description": "Open Source DeepWiki: AI-Powered Wiki Generator for GitHub/Gitlab/Bitbucket Repositories. Join the discord: https://discord.gg/gMwThUMeme", + "default_branch": "main", + "archived": false, + "fork": false, + "pushed_at": "2026-09-03T20:29:26Z", + "license_metadata": "MIT", + "topics": [ + "ai", + "codex", + "gemini", + "github", + "grok-cli", + "ollama", + "open-source", + "openai", + "openrouter", + "self-hosted", + "wiki" + ] + }, + { + "repository": "aaif-goose/goose", + "url": "https://github.com/aaif-goose/goose", + "description": "an open source, extensible AI agent that goes beyond code suggestions - install, execute, edit, and test with any LLM", + "default_branch": "main", + "archived": false, + "fork": false, + "pushed_at": "2026-09-19T18:33:03Z", + "license_metadata": "Apache-2.0", + "topics": [ + "acp", + "ai", + "ai-agents", + "mcp" + ] + }, + { + "repository": "thedotmack/claude-mem", + "url": "https://github.com/thedotmack/claude-mem", + "description": "Persistent Context Across Sessions for Every Agent \u2013 Captures everything your agent does during sessions, compresses it with AI, and injects relevant context back into future sessions. Works with Claude Code, OpenClaw, Codex, Gemini, Hermes, Copilot, OpenCode + More", + "default_branch": "main", + "archived": false, + "fork": false, + "pushed_at": "2026-09-19T18:14:21Z", + "license_metadata": "Apache-2.0", + "topics": [ + "ai", + "ai-agents", + "ai-memory", + "anthropic", + "artificial-intelligence", + "chromadb", + "claude", + "claude-agent-sdk", + "claude-agents", + "claude-code", + "claude-code-plugin", + "claude-skills", + "embeddings", + "long-term-memory", + "mem0", + "memory-engine", + "openmemory", + "rag", + "sqlite", + "supermemory" + ] + }, + { + "repository": "coleam00/Archon", + "url": "https://github.com/coleam00/Archon", + "description": "The first open-source harness builder for AI coding. Make AI coding deterministic and repeatable.", + "default_branch": "dev", + "archived": false, + "fork": false, + "pushed_at": "2026-09-18T12:55:28Z", + "license_metadata": "MIT", + "topics": [ + "ai", + "automation", + "bun", + "claude", + "cli", + "coding-assistant", + "developer-tools", + "typescript", + "workflow-engine", + "yaml" + ] + }, + { + "repository": "shiyu-coder/Kronos", + "url": "https://github.com/shiyu-coder/Kronos", + "description": "Kronos: A Foundation Model for the Language of Financial Markets", + "default_branch": "master", + "archived": false, + "fork": false, + "pushed_at": "2026-04-13T12:38:49Z", + "license_metadata": "MIT", + "topics": [] + }, + { + "repository": "mem0ai/mem0", + "url": "https://github.com/mem0ai/mem0", + "description": "The Memory Layer for AI Agents - Drop-in memory infrastructure for AI agents and apps. Context that persists. Built for production.", + "default_branch": "main", + "archived": false, + "fork": false, + "pushed_at": "2026-09-19T01:14:39Z", + "license_metadata": "Apache-2.0", + "topics": [ + "agentic-memory", + "agentic-memory-system", + "agents", + "ai", + "ai-agents", + "chatgpt", + "genai", + "llm", + "long-term-memory", + "memory", + "memory-management", + "python", + "rag", + "state-management" + ] + }, + { + "repository": "Tracer-Cloud/opensre", + "url": "https://github.com/Tracer-Cloud/opensre", + "description": "Build your own AI SRE agents. The open source toolkit for the AI era.", + "default_branch": "main", + "archived": false, + "fork": false, + "pushed_at": "2026-09-20T01:21:37Z", + "license_metadata": "Apache-2.0", + "topics": [ + "ai-sre", + "alerting", + "datadog", + "grafana", + "incident-management", + "observability", + "remediation", + "root-cause-analysis", + "site-reliability-engineering", + "slack", + "sre" + ] + }, + { + "repository": "garrytan/gbrain", + "url": "https://github.com/garrytan/gbrain", + "description": "Garry's Opinionated OpenClaw/Hermes Agent Brain", + "default_branch": "master", + "archived": false, + "fork": false, + "pushed_at": "2026-09-17T21:00:26Z", + "license_metadata": "MIT", + "topics": [] + }, + { + "repository": "MemPalace/mempalace", + "url": "https://github.com/MemPalace/mempalace", + "description": "The best-benchmarked open-source AI memory system. And it's free.", + "default_branch": "develop", + "archived": false, + "fork": false, + "pushed_at": "2026-09-18T07:55:23Z", + "license_metadata": "MIT", + "topics": [ + "ai", + "chromadb", + "llm", + "mcp", + "memory", + "python" + ] + }, + { + "repository": "vinta/awesome-python", + "url": "https://github.com/vinta/awesome-python", + "description": "The definitive list that answers \"I want to do X in Python, which tool should I use?\"", + "default_branch": "master", + "archived": false, + "fork": false, + "pushed_at": "2026-09-18T22:11:03Z", + "license_metadata": "NOASSERTION", + "topics": [ + "awesome", + "awesome-list", + "python", + "python-frameworks", + "python-libraries", + "python-tools" + ] + }, + { + "repository": "alirezarezvani/claude-skills", + "url": "https://github.com/alirezarezvani/claude-skills", + "description": "380 Claude Code skills & agent skills & plugins (30+ Agents, 70+ custom commands, 380+ skills, customizable references, scripts)for Claude Code, Codex, Gemini CLI, Cursor, and 8 more coding agents \u2014 engineering, marketing, product, compliance, C-level advisory, research, business operations, commercial & finance, and your daily productivity skills.", + "default_branch": "main", + "archived": false, + "fork": false, + "pushed_at": "2026-08-30T09:46:16Z", + "license_metadata": "MIT", + "topics": [ + "agent-plugins", + "agent-skills", + "agentic-ai", + "ai-coding-agent", + "anthropic-claude", + "claude-ai", + "claude-code", + "claude-code-plugins", + "claude-code-skills", + "claude-skills", + "codex-skills", + "coding-agent-plugins", + "cursor-skills", + "developer-tools", + "gemini-cli-skills", + "openai-codex", + "openclaw", + "openclaw-plugins", + "openclaw-skills", + "prompt-engineering" + ] + }, + { + "repository": "unslothai/unsloth", + "url": "https://github.com/unslothai/unsloth", + "description": "Local UI to run and train LLMs and diffusion models. Supports GGUF, MLX, Qwen3.8, DeepSeek-V4, MiniMax-H3, Gemma 4, FLUX and more.", + "default_branch": "main", + "archived": false, + "fork": false, + "pushed_at": "2026-09-20T01:43:23Z", + "license_metadata": "Apache-2.0", + "topics": [ + "agent", + "ai", + "chatgpt", + "deepseek", + "fine-tuning", + "gemma", + "image-generation", + "llama", + "llm", + "llms", + "openai", + "python", + "qwen", + "reinforcement-learning", + "self-hosted", + "stable-diffusion", + "text-to-speech", + "tts", + "ui", + "unsloth" + ] + }, + { + "repository": "hesreallyhim/awesome-claude-code", + "url": "https://github.com/hesreallyhim/awesome-claude-code", + "description": "A hand-picked collection of the finest of resources for the most awesome of agents, Claude Code, the undisputed champion of coding companions, from the unstoppable team at Anthropic PBC. A delectable showcase of top tier skills, ambidextrous agents, scintillating status lines, top notch developer tooling, and also we have plugins", + "default_branch": "main", + "archived": false, + "fork": false, + "pushed_at": "2026-09-20T00:29:54Z", + "license_metadata": "NOASSERTION", + "topics": [ + "agent-skills", + "agentic-code", + "agentic-coding", + "ai-workflow-optimization", + "ai-workflows", + "anthropic", + "anthropic-claude", + "awesome", + "awesome-claude-code", + "awesome-list", + "awesome-lists", + "awesome-resources", + "claude", + "claude-code", + "coding-agent", + "coding-agents", + "coding-assistant", + "coding-assistants", + "llm" + ] + }, + { + "repository": "mattpocock/skills", + "url": "https://github.com/mattpocock/skills", + "description": "Skills for Real Engineers. Straight from my .agents directory.", + "default_branch": "main", + "archived": false, + "fork": false, + "pushed_at": "2026-09-18T10:12:48Z", + "license_metadata": "MIT", + "topics": [] + }, + { + "repository": "TauricResearch/TradingAgents", + "url": "https://github.com/TauricResearch/TradingAgents", + "description": "TradingAgents: Multi-Agents LLM Financial Trading Framework", + "default_branch": "main", + "archived": false, + "fork": false, + "pushed_at": "2026-09-18T05:43:45Z", + "license_metadata": "Apache-2.0", + "topics": [ + "agent", + "finance", + "llm", + "multiagent", + "trading" + ] + }, + { + "repository": "onyx-dot-app/onyx", + "url": "https://github.com/onyx-dot-app/onyx", + "description": "Open Source AI Platform - AI Chat with advanced features that works with every LLM", + "default_branch": "main", + "archived": false, + "fork": false, + "pushed_at": "2026-09-19T10:13:40Z", + "license_metadata": "NOASSERTION", + "topics": [ + "ai", + "ai-chat", + "chatgpt", + "chatui", + "enterprise-search", + "gen-ai", + "information-retrieval", + "llm", + "llm-ui", + "nextjs", + "python", + "rag", + "self-hosted", + "vector-search" + ] + }, + { + "repository": "luongnv89/claude-howto", + "url": "https://github.com/luongnv89/claude-howto", + "description": "A visual, example-driven guide to Claude Code \u2014 from basic concepts to advanced agents, with copy-paste templates that bring immediate value.", + "default_branch": "main", + "archived": false, + "fork": false, + "pushed_at": "2026-09-19T23:15:37Z", + "license_metadata": "MIT", + "topics": [ + "claude-code", + "guide", + "tutorial" + ] + }, + { + "repository": "google-ai-edge/gallery", + "url": "https://github.com/google-ai-edge/gallery", + "description": "A gallery that showcases on-device ML/GenAI use cases and allows people to try and use models locally.", + "default_branch": "main", + "archived": false, + "fork": false, + "pushed_at": "2026-09-18T21:55:53Z", + "license_metadata": "Apache-2.0", + "topics": [] + }, + { + "repository": "google-ai-edge/LiteRT-LM", + "url": "https://github.com/google-ai-edge/LiteRT-LM", + "description": "LiteRT-LM is Google's production-ready, high-performance, open-source inference framework for deploying Large Language Models on edge devices.", + "default_branch": "main", + "archived": false, + "fork": false, + "pushed_at": "2026-09-19T22:59:03Z", + "license_metadata": "Apache-2.0", + "topics": [ + "edge-ai", + "on-device-ai", + "on-device-llm" + ] + }, + { + "repository": "YishenTu/claudian", + "url": "https://github.com/YishenTu/claudian", + "description": "An Obsidian plugin that embeds Claude Code/Codex as an AI collaborator in your vault", + "default_branch": "main", + "archived": false, + "fork": false, + "pushed_at": "2026-09-20T01:59:51Z", + "license_metadata": "MIT", + "topics": [ + "claude-code", + "codex", + "ide", + "obsidian", + "obsidian-plugin", + "productivity" + ] + }, + { + "repository": "multica-ai/andrej-karpathy-skills", + "url": "https://github.com/multica-ai/andrej-karpathy-skills", + "description": "A single CLAUDE.md file to improve Claude Code behavior, derived from Andrej Karpathy's observations on LLM coding pitfalls.", + "default_branch": "main", + "archived": false, + "fork": false, + "pushed_at": "2026-04-20T10:05:04Z", + "license_metadata": null, + "topics": [] + }, + { + "repository": "Shubhamsaboo/awesome-llm-apps", + "url": "https://github.com/Shubhamsaboo/awesome-llm-apps", + "description": "100+ AI Agents, Agent Skills and RAG Apps - Free and Open Source.", + "default_branch": "main", + "archived": false, + "fork": false, + "pushed_at": "2026-09-19T07:54:25Z", + "license_metadata": "Apache-2.0", + "topics": [ + "agents", + "llms", + "python", + "rag" + ] + }, + { + "repository": "HKUDS/OpenHarness", + "url": "https://github.com/HKUDS/OpenHarness", + "description": "\"OpenHarness: Open Agent Harness with a Built-in Personal Agent--Ohmo!\"", + "default_branch": "main", + "archived": false, + "fork": false, + "pushed_at": "2026-06-04T02:42:40Z", + "license_metadata": "MIT", + "topics": [] + }, + { + "repository": "quemsah/awesome-claude-plugins", + "url": "https://github.com/quemsah/awesome-claude-plugins", + "description": "Automated collection of Claude Code plugin adoption metrics across GitHub repositories using n8n workflows", + "default_branch": "main", + "archived": false, + "fork": false, + "pushed_at": "2026-09-19T07:37:46Z", + "license_metadata": null, + "topics": [ + "awesome-list", + "claude-code", + "claude-code-plugin", + "claude-code-plugins", + "claude-code-plugins-marketplace" + ] + }, + { + "repository": "karpathy/autoresearch", + "url": "https://github.com/karpathy/autoresearch", + "description": "AI agents running research on single-GPU nanochat training automatically", + "default_branch": "master", + "archived": false, + "fork": false, + "pushed_at": "2026-03-26T00:07:37Z", + "license_metadata": null, + "topics": [] + }, + { + "repository": "shanraisshan/claude-code-best-practice", + "url": "https://github.com/shanraisshan/claude-code-best-practice", + "description": "from vibe coding to agentic engineering - practice makes claude perfect", + "default_branch": "main", + "archived": false, + "fork": false, + "pushed_at": "2026-09-19T06:40:19Z", + "license_metadata": "MIT", + "topics": [ + "agentic-ai", + "agentic-coding", + "agentic-engineering", + "agentic-workflow", + "ai", + "ai-agents", + "anthropic", + "best-practices", + "boris", + "claude", + "claude-ai", + "claude-code", + "claude-code-agents", + "claude-code-best-practices", + "claude-code-commands", + "claude-code-skills", + "context-engineering", + "pakistan", + "pakistani-developer", + "vibe-coding" + ] + }, + { + "repository": "abhigyanpatwari/GitNexus", + "url": "https://github.com/abhigyanpatwari/GitNexus", + "description": "GitNexus: The Zero-Server Code Intelligence Engine ", + "default_branch": "main", + "archived": false, + "fork": false, + "pushed_at": "2026-09-19T15:09:39Z", + "license_metadata": "NOASSERTION", + "topics": [] + }, + { + "repository": "vectorize-io/hindsight", + "url": "https://github.com/vectorize-io/hindsight", + "description": "Hindsight: Agent Memory That Learns", + "default_branch": "main", + "archived": false, + "fork": false, + "pushed_at": "2026-09-19T14:30:38Z", + "license_metadata": "MIT", + "topics": [ + "agentic-ai", + "agents", + "ai-memory", + "memory" + ] + }, + { + "repository": "wanshuiyin/Auto-claude-code-research-in-sleep", + "url": "https://github.com/wanshuiyin/Auto-claude-code-research-in-sleep", + "description": "ARIS \u2694\ufe0f (Auto-Research-In-Sleep) \u2014 Lightweight Markdown-only skills for autonomous ML research: cross-model review loops, idea discovery, and experiment automation. No framework, no lock-in \u2014 works with Claude Code, Codex, OpenClaw, or any LLM agent.", + "default_branch": "main", + "archived": false, + "fork": false, + "pushed_at": "2026-09-18T18:05:10Z", + "license_metadata": "MIT", + "topics": [ + "ai-research", + "ai-tools", + "aris", + "autonomous-agent", + "claude", + "claude-code", + "claude-code-skills", + "codex", + "deep-learning", + "gpt", + "idea-generation", + "llm", + "machine-learning", + "mcp", + "mcp-server", + "ml-research", + "openai", + "paper-review", + "paper-writing", + "research-automation" + ] + }, + { + "repository": "wshobson/agents", + "url": "https://github.com/wshobson/agents", + "description": "Multi-harness agentic plugin marketplace for Claude Code, Codex, Cursor, OpenCode, GitHub Copilot, Google Antigravity, and Pi", + "default_branch": "main", + "archived": false, + "fork": false, + "pushed_at": "2026-09-19T02:36:23Z", + "license_metadata": "MIT", + "topics": [ + "agent-skills", + "agentic-ai", + "ai-agents", + "anthropic", + "antigravity", + "claude", + "claude-code", + "claude-code-marketplace", + "claude-code-plugin", + "claude-skills", + "codex", + "coding-agents", + "cursor", + "cursor-rules", + "github-copilot", + "mcp", + "multi-agent", + "opencode", + "pi-coding-agent", + "subagents" + ] + }, + { + "repository": "EverMind-AI/EverOS", + "url": "https://github.com/EverMind-AI/EverOS", + "description": "One portable memory layer for every AI agent: local-first, Markdown-native, user-owned, and self-evolving across apps, tools, and workflows.", + "default_branch": "main", + "archived": false, + "fork": false, + "pushed_at": "2026-09-09T02:59:13Z", + "license_metadata": "Apache-2.0", + "topics": [ + "agent-memory", + "agentic-ai", + "ai", + "chats", + "clawdbot", + "clawdbot-skill", + "deepseek-harness", + "dsh", + "dsh-plugin", + "llm", + "long-term-memory", + "mcp", + "memory", + "memory-management", + "python3", + "rag", + "skills" + ] + }, + { + "repository": "promptfoo/promptfoo", + "url": "https://github.com/promptfoo/promptfoo", + "description": "Test your prompts, agents, and RAGs. Red teaming/pentesting/vulnerability scanning for AI. Compare performance of GPT, Claude, Gemini, DeepSeek, and more. Simple declarative configs with command line and CI/CD integration. Used by OpenAI and Anthropic.", + "default_branch": "main", + "archived": false, + "fork": false, + "pushed_at": "2026-09-20T01:37:28Z", + "license_metadata": "MIT", + "topics": [ + "ci", + "ci-cd", + "cicd", + "evaluation", + "evaluation-framework", + "llm", + "llm-eval", + "llm-evaluation", + "llm-evaluation-framework", + "llmops", + "pentesting", + "prompt-engineering", + "prompt-testing", + "prompts", + "rag", + "red-teaming", + "testing", + "vulnerability-scanners" + ] + }, + { + "repository": "langchain-ai/deepagents", + "url": "https://github.com/langchain-ai/deepagents", + "description": "The batteries-included agent harness.", + "default_branch": "main", + "archived": false, + "fork": false, + "pushed_at": "2026-09-19T22:51:43Z", + "license_metadata": "MIT", + "topics": [ + "ai", + "deepagents", + "harness", + "harness-engineering", + "langchain", + "langgraph", + "python", + "typescript" + ] + }, + { + "repository": "NousResearch/hermes-agent", + "url": "https://github.com/NousResearch/hermes-agent", + "description": "The agent that grows with you", + "default_branch": "main", + "archived": false, + "fork": false, + "pushed_at": "2026-09-20T01:51:37Z", + "license_metadata": "MIT", + "topics": [ + "ai", + "ai-agent", + "ai-agents", + "anthropic", + "chatgpt", + "claude", + "claude-code", + "codex", + "hermes", + "hermes-agent", + "llm", + "nous-research", + "openai" + ] + }, + { + "repository": "pbakaus/impeccable", + "url": "https://github.com/pbakaus/impeccable", + "description": "The design language that makes your AI harness better at design.", + "default_branch": "main", + "archived": false, + "fork": false, + "pushed_at": "2026-09-19T01:09:27Z", + "license_metadata": "Apache-2.0", + "topics": [] + }, + { + "repository": "andrewyng/context-hub", + "url": "https://github.com/andrewyng/context-hub", + "description": null, + "default_branch": "main", + "archived": false, + "fork": false, + "pushed_at": "2026-05-31T18:41:45Z", + "license_metadata": "MIT", + "topics": [] + }, + { + "repository": "msitarzewski/agency-agents", + "url": "https://github.com/msitarzewski/agency-agents", + "description": "A complete AI agency at your fingertips - From frontend wizards to Reddit community ninjas, from whimsy injectors to reality checkers. Each agent is a specialized expert with personality, processes, and proven deliverables.", + "default_branch": "main", + "archived": false, + "fork": false, + "pushed_at": "2026-09-12T16:57:21Z", + "license_metadata": "MIT", + "topics": [] + }, + { + "repository": "lightpanda-io/browser", + "url": "https://github.com/lightpanda-io/browser", + "description": "Lightpanda: the headless browser designed for AI and automation", + "default_branch": "main", + "archived": false, + "fork": false, + "pushed_at": "2026-09-20T01:04:25Z", + "license_metadata": "AGPL-3.0", + "topics": [ + "browser", + "browser-automation", + "cdp", + "headless", + "lightpanda", + "playwright", + "puppeteer", + "zig" + ] + }, + { + "repository": "666ghj/MiroFish", + "url": "https://github.com/666ghj/MiroFish", + "description": "A Simple and Universal Swarm Intelligence Engine, Predicting Anything. \u7b80\u6d01\u901a\u7528\u7684\u7fa4\u4f53\u667a\u80fd\u5f15\u64ce\uff0c\u9884\u6d4b\u4e07\u7269", + "default_branch": "main", + "archived": false, + "fork": false, + "pushed_at": "2026-09-16T03:31:58Z", + "license_metadata": "AGPL-3.0", + "topics": [ + "agent-memory", + "financial-forecasting", + "future-prediction", + "knowledge-graph", + "llms", + "multi-agent-simulation", + "public-opinion-analysis", + "python3", + "social-prediction", + "swarm-intelligence" + ] + }, + { + "repository": "shareAI-lab/learn-claude-code", + "url": "https://github.com/shareAI-lab/learn-claude-code", + "description": "Bash is all you need - A nano claude code\u2013like \u300cagent harness\u300d, built from 0 to 1", + "default_branch": "main", + "archived": false, + "fork": false, + "pushed_at": "2026-08-26T16:38:22Z", + "license_metadata": "MIT", + "topics": [ + "agent", + "agent-development", + "ai-agent", + "claude", + "claude-code", + "educational", + "llm", + "python", + "teaching", + "tutorial" + ] + }, + { + "repository": "obra/superpowers", + "url": "https://github.com/obra/superpowers", + "description": "An agentic skills framework & software development methodology that works.", + "default_branch": "main", + "archived": false, + "fork": false, + "pushed_at": "2026-09-19T00:32:57Z", + "license_metadata": "MIT", + "topics": [ + "ai", + "brainstorming", + "coding", + "obra", + "sdlc", + "skills", + "subagent-driven-development", + "superpowers" + ] + }, + { + "repository": "langchain-ai/open-swe", + "url": "https://github.com/langchain-ai/open-swe", + "description": "An Open-Source Asynchronous Coding Agent", + "default_branch": "main", + "archived": false, + "fork": false, + "pushed_at": "2026-09-20T01:23:20Z", + "license_metadata": "MIT", + "topics": [ + "agent", + "agents", + "ai", + "anthropic", + "claudecode", + "llm", + "llms", + "openai" + ] + }, + { + "repository": "jarrodwatts/claude-hud", + "url": "https://github.com/jarrodwatts/claude-hud", + "description": "A Claude Code plugin that shows what's happening - context usage, active tools, running agents, and todo progress", + "default_branch": "main", + "archived": false, + "fork": false, + "pushed_at": "2026-09-19T21:32:54Z", + "license_metadata": "MIT", + "topics": [ + "anthropic", + "claude", + "claude-code", + "cli", + "plugin", + "statusline", + "typescript" + ] + }, + { + "repository": "koala73/worldmonitor", + "url": "https://github.com/koala73/worldmonitor", + "description": "Real-time global intelligence dashboard. AI-powered news aggregation, geopolitical monitoring, and infrastructure tracking in a unified situational awareness interface", + "default_branch": "main", + "archived": false, + "fork": false, + "pushed_at": "2026-09-19T20:34:16Z", + "license_metadata": "AGPL-3.0", + "topics": [ + "agent", + "ai", + "dashboard", + "geopolitics", + "mcp", + "mcp-server", + "monitoring", + "news", + "opensource", + "osint", + "palantir", + "situation" + ] + }, + { + "repository": "VectifyAI/PageIndex", + "url": "https://github.com/VectifyAI/PageIndex", + "description": "\ud83d\udcd1 PageIndex: Document Index for Vectorless, Reasoning-based RAG", + "default_branch": "main", + "archived": false, + "fork": false, + "pushed_at": "2026-09-19T13:53:40Z", + "license_metadata": "MIT", + "topics": [ + "agentic-ai", + "agents", + "ai", + "ai-agents", + "context-engineering", + "information-retrieval", + "llm", + "rag", + "reasoning", + "retrieval", + "retrieval-augmented-generation", + "vector-database" + ] + }, + { + "repository": "ruvnet/ruflo", + "url": "https://github.com/ruvnet/ruflo", + "description": "\ud83c\udf0a The original agent harness. Deploy intelligent multi-player swarms, coordinate autonomous workflows, and build conversational AI systems. Features adaptive memory, self-learning intelligence, federation, vector RAG integration, and native Claude Code / Codex / Hermes and many more Integrated", + "default_branch": "main", + "archived": false, + "fork": false, + "pushed_at": "2026-09-19T06:23:46Z", + "license_metadata": "MIT", + "topics": [ + "agentic-ai", + "agentic-framework", + "agentic-workflow", + "agents", + "ai-agents", + "ai-assistant", + "ai-skills", + "autonomous-agents", + "claude-code", + "codex", + "dsh-plugin", + "harness", + "mcp-server", + "multi-agent", + "multi-agent-systems", + "npm", + "skills", + "swarm", + "swarm-intelligence", + "typescript" + ] + }, + { + "repository": "comet-ml/opik", + "url": "https://github.com/comet-ml/opik", + "description": "Debug, evaluate, and monitor your LLM applications, RAG systems, and agentic workflows with comprehensive tracing, automated evaluations, and production-ready dashboards.", + "default_branch": "main", + "archived": false, + "fork": false, + "pushed_at": "2026-09-19T09:37:41Z", + "license_metadata": "Apache-2.0", + "topics": [ + "evaluation", + "hacktoberfest", + "hacktoberfest2025", + "langchain", + "llama-index", + "llm", + "llm-evaluation", + "llm-observability", + "llmops", + "open-source", + "openai", + "playground", + "prompt-engineering" + ] + }, + { + "repository": "affaan-m/ECC", + "url": "https://github.com/affaan-m/ECC", + "description": "The agent harness performance optimization system. Skills, instincts, memory, security, and research-first development for Claude Code, Codex, Opencode, Cursor and beyond.", + "default_branch": "main", + "archived": false, + "fork": false, + "pushed_at": "2026-09-20T00:01:14Z", + "license_metadata": "MIT", + "topics": [ + "ai-agents", + "anthropic", + "claude", + "claude-code", + "developer-tools", + "llm", + "mcp", + "productivity" + ] + } + ] +} diff --git a/catalogs/convergence-practice/source-review.json b/catalogs/convergence-practice/source-review.json new file mode 100644 index 000000000..20756f9a2 --- /dev/null +++ b/catalogs/convergence-practice/source-review.json @@ -0,0 +1,810 @@ +{ + "schema_version": 1, + "kind": "bounded_public_convergence_source_review", + "retrieved_at": "2026-09-20T02:07:12.602467+00:00", + "scope": "One bounded source-review wave spanning evaluation, recovery, local inference and context. This is not a universal SOTA ranking or an installation/runtime acceptance.", + "public_base": { + "repository": "seathatflowsinourveins/native-agent-stack", + "commit": "bf99d340f798bf72cc3104619edc482882eb423d", + "decision_index_url": "https://github.com/seathatflowsinourveins/native-agent-stack/blob/bf99d340f798bf72cc3104619edc482882eb423d/catalogs/us-equities/decision-index.json", + "decision_index_sha256": "7c6b0bb779cdbc46077e63d545f68219521a661343e48fdc3dd19a8c2534f593", + "repository_count": 501 + }, + "public_inventory": { + "owned_count": 1, + "starred_count": 342, + "starred_pagination_pages": 4, + "owned_endpoint": "https://api.github.com/users/seathatflowsinourveins/repos?type=owner&per_page=100", + "starred_endpoint": "https://api.github.com/users/seathatflowsinourveins/starred?per_page=100", + "retrieved_at": "2026-09-20T02:00:50.538618+00:00", + "inventory_files": [ + "public-owned.json", + "public-starred.json" + ], + "missing_public_stars_from_index": [], + "nonpublic_records_retained": 0, + "fields_allowlisted": true, + "star_count_used_as_quality_rank": false + }, + "awesome_sources": [ + { + "repository": "awesome-foss/awesome-sysadmin", + "source_commit": "a0f64294fa6c672767f4dffc978023194bf99550", + "retrieved_at": "2026-09-20T02:01:19.828227+00:00", + "readme_url": "https://github.com/awesome-foss/awesome-sysadmin/blob/a0f64294fa6c672767f4dffc978023194bf99550/README.md", + "readme_blob_sha": "bbd268f26f3284d47d07461430a7185d86e085bd", + "discovery_only": true, + "license_sources": [ + { + "path": "LICENSE.txt", + "url": "https://github.com/awesome-foss/awesome-sysadmin/blob/a0f64294fa6c672767f4dffc978023194bf99550/LICENSE.txt", + "blob_sha": "92de3a9259139719271fbcebd47e775b7309c900", + "sha256": "2ee4926edd20442e2237f93f90bcb519f97b4200f0334426e76f273be6a764d4", + "bytes": 14494 + } + ], + "bytes": 76102, + "sha256": "c86e591abee4eed1f47675c81f2cfc1b07fa3f568bb2e54f280b508e8784849b", + "license_at_pin": "CC-BY-SA-4.0", + "actual_outbound_links": [ + { + "repository": "restic/restic", + "target_url": "https://github.com/restic/restic", + "source_url": "https://github.com/awesome-foss/awesome-sysadmin/blob/a0f64294fa6c672767f4dffc978023194bf99550/README.md#L111", + "disposition": "selected primary review" + }, + { + "repository": "borgbackup/borg", + "target_url": "https://github.com/borgbackup/borg", + "source_url": "https://github.com/awesome-foss/awesome-sysadmin/blob/a0f64294fa6c672767f4dffc978023194bf99550/README.md#L99", + "disposition": "discovery only; alternate backup engine overlaps retained restic" + }, + { + "repository": "garethgeorge/backrest", + "target_url": "https://github.com/garethgeorge/backrest", + "source_url": "https://github.com/awesome-foss/awesome-sysadmin/blob/a0f64294fa6c672767f4dffc978023194bf99550/README.md#L95", + "disposition": "discovery only; GUI/orchestrator adds another service before recovery acceptance" + } + ], + "scope": "Selected matching entries reviewed; the full list is not audited and does not establish tool quality or compatibility." + }, + { + "repository": "punkpeye/awesome-mcp-servers", + "source_commit": "393b4e9fafb0348e5a1c2a4ef5a8719b0d85e061", + "retrieved_at": "2026-09-20T02:01:19.693991+00:00", + "readme_url": "https://github.com/punkpeye/awesome-mcp-servers/blob/393b4e9fafb0348e5a1c2a4ef5a8719b0d85e061/README.md", + "readme_blob_sha": "45f85250febd3da224f78340c775ff50d72842fb", + "discovery_only": true, + "bytes": 1756337, + "sha256": "9b6a1ee410fde8132c92444f15cc153c2d6c4866dd74c4ac8b428981d49f777c", + "license_sources": [ + { + "path": "LICENSE", + "url": "https://github.com/punkpeye/awesome-mcp-servers/blob/393b4e9fafb0348e5a1c2a4ef5a8719b0d85e061/LICENSE", + "blob_sha": "7f5a6e30ea24860b1528cfb7a54cf777650fd46a", + "sha256": "8cb2c0a4c6590502ffb793005c0f1b440ca4db34dcc209bbe9cf34d3b52a7e07", + "bytes": 1125 + } + ], + "license_at_pin": "MIT", + "actual_outbound_links": [ + { + "repository": "oraios/serena", + "target_url": "https://github.com/oraios/serena", + "source_url": "https://github.com/punkpeye/awesome-mcp-servers/blob/393b4e9fafb0348e5a1c2a4ef5a8719b0d85e061/README.md#L807", + "disposition": "discovery only; already cataloged symbolic-edit option, require concrete gap" + }, + { + "repository": "jagoff/memo", + "target_url": "https://github.com/jagoff/memo", + "source_url": "https://github.com/punkpeye/awesome-mcp-servers/blob/393b4e9fafb0348e5a1c2a4ef5a8719b0d85e061/README.md#L2785", + "disposition": "discovery only; memory/MLX feature list overlaps scoped memory and retrieval; no adoption" + } + ], + "scope": "Selected matching entries reviewed; the full list is not audited and does not establish tool quality or compatibility." + }, + { + "repository": "e2b-dev/awesome-ai-agents", + "source_commit": "9596ab1e69fbbb6c95291141c8dab5c181083216", + "retrieved_at": "2026-09-20T02:01:19.855817+00:00", + "readme_url": "https://github.com/e2b-dev/awesome-ai-agents/blob/9596ab1e69fbbb6c95291141c8dab5c181083216/README.md", + "readme_blob_sha": "b701937e2c389d3e9fe3830617e4c177e37c06fe", + "discovery_only": true, + "license_sources": [ + { + "path": "LICENSE.md", + "url": "https://github.com/e2b-dev/awesome-ai-agents/blob/9596ab1e69fbbb6c95291141c8dab5c181083216/LICENSE.md", + "blob_sha": "d0e1a6978bcaf61741acd6f3295bdd5aa348cc29", + "sha256": "7a31126d7adf334ccdcbefd35c2fe98254db746113c6b80e9943b1800d36fc39", + "bytes": 19089 + } + ], + "bytes": 212139, + "sha256": "db062c798654c5f6dec1ddedf7ab823c70d11d2f66ccf2a7036171f1e5b6eaa8", + "license_at_pin": "CC-BY-NC-SA-4.0", + "actual_outbound_links": [ + { + "repository": "princeton-nlp/SWE-agent", + "target_url": "https://github.com/princeton-nlp/SWE-agent", + "source_url": "https://github.com/e2b-dev/awesome-ai-agents/blob/9596ab1e69fbbb6c95291141c8dab5c181083216/README.md#L2577", + "disposition": "discovery only; another coding-agent runtime; historical score claims were not accepted as evaluation evidence" + } + ], + "scope": "Selected matching entries reviewed; the full list is not audited and does not establish tool quality or compatibility." + } + ], + "candidates": [ + { + "repository": "eyuansu62/agent-retrieval-bench", + "repository_key": "eyuansu62/agent-retrieval-bench", + "version": "v0.2.1", + "published_package_version": null, + "source_commit": "b487f3866cc13dd971819cb902517a6a50282404", + "retrieved_at": "2026-09-20T02:02:08.467015+00:00", + "release_url": "https://github.com/eyuansu62/agent-retrieval-bench/releases/tag/v0.2.1", + "published_at": "2026-07-26T01:08:48Z", + "source_depth": "pinned_readme_license_and_selected_supporting_source", + "license_at_pin": { + "spdx": "MIT", + "url": "https://github.com/eyuansu62/agent-retrieval-bench/blob/b487f3866cc13dd971819cb902517a6a50282404/LICENSE", + "sha256": "f821747bf0011218a8d18a6030b3343384a2c68478e643e26e057732fc89aa31" + }, + "catalog_coverage": { + "repository": "eyuansu62/agent-retrieval-bench", + "in_public_decision_index": true, + "public_references": [ + { + "kind": "catalog_card", + "path": "catalogs/us-equities/foundation-memory.json", + "pointer": "/entries/33", + "id": "foundation-agent-retrieval-bench", + "decision": "conditional", + "evidence_level": "source_review" + } + ], + "public_star": false, + "in_maintained_registry": true, + "maintained_selection": "trial", + "maintained_pin": "v0.2.1" + }, + "discovery_provenance": [ + { + "kind": "public_owned_repository_catalog", + "source_url": "https://github.com/seathatflowsinourveins/native-agent-stack/blob/bf99d340f798bf72cc3104619edc482882eb423d/catalogs/us-equities/decision-index.json" + } + ], + "sources": [ + { + "path": "LICENSE", + "url": "https://github.com/eyuansu62/agent-retrieval-bench/blob/b487f3866cc13dd971819cb902517a6a50282404/LICENSE", + "blob_sha": "8d60c0515b80b0e50dc5877f209aa84425106b82", + "sha256": "f821747bf0011218a8d18a6030b3343384a2c68478e643e26e057732fc89aa31", + "bytes": 1091 + }, + { + "path": "README.md", + "url": "https://github.com/eyuansu62/agent-retrieval-bench/blob/b487f3866cc13dd971819cb902517a6a50282404/README.md", + "blob_sha": "d5be7aaea6e17ace0a36b319c03e71199aa84ef6", + "sha256": "494b73638318f3e8958422e6ac49e6c62d13973c0e4d947c186ed3b305c3dbc4", + "bytes": 9985 + }, + { + "path": "pyproject.toml", + "url": "https://github.com/eyuansu62/agent-retrieval-bench/blob/b487f3866cc13dd971819cb902517a6a50282404/pyproject.toml", + "blob_sha": "a5ec148c61f2449edb85e415d4218d2cbbba35d7", + "sha256": "05889ec96609b4d5b557ac21b9894a3d308930dbe7d8d11deb130abdc9d2e317", + "bytes": 729 + }, + { + "path": "DATA_LICENSE.md", + "url": "https://github.com/eyuansu62/agent-retrieval-bench/blob/b487f3866cc13dd971819cb902517a6a50282404/DATA_LICENSE.md", + "blob_sha": "ec257b002b58516a673e747c7fe4be69bf49ee17", + "sha256": "c503889e22cb019a92ddc8c51e704e89c2e85f1797d5f97cc3c022dd49f1a19e", + "bytes": 1404 + } + ], + "runtime_executed": false, + "lane": "evaluation/context", + "decision": "investigate", + "why": "Adds frozen pre-change file-retrieval tasks, token-budgeted context yield and abstention cases to the existing small repository-specific evaluation.", + "overlap": "Evaluates retrieval rather than adding a runtime. Complement existing rg, ast-grep, SocratiCode and selected document retrieval.", + "compatibility": "Python >=3.10. Corpus snapshots retain each upstream project license; evaluator/metadata/documentation are MIT.", + "acceptance": "Freeze a stratified public subset and hashes before runs. Compare lexical, existing semantic and selective-abstention lanes under the same all_files candidates and 4k/8k token budgets. Report Recall@k, MRR, BCY, no-gold errors, time and actual provider usage separately. Add independent edit/test tasks before claiming patch improvement.", + "limits": [ + "File hits do not prove correct spans or test-passing edits.", + "Published method rankings are workload-specific; no universal best embedding claim.", + "No corpus download, model call or benchmark execution in this audit." + ], + "review_level": "source_review", + "evidence_level": "pinned_source_review", + "source_pin_distinction": { + "reviewed_release_tag": "v0.2.1", + "reviewed_release_commit": "b487f3866cc13dd971819cb902517a6a50282404", + "maintained_prior_source_commit": "07014c986f3deadb1548c62b32c0ffbe6a81465d", + "maintained_source_url": "https://github.com/eyuansu62/agent-retrieval-bench/blob/07014c986f3deadb1548c62b32c0ffbe6a81465d/README.md", + "comparison_url": "https://github.com/eyuansu62/agent-retrieval-bench/compare/b487f3866cc13dd971819cb902517a6a50282404...07014c986f3deadb1548c62b32c0ffbe6a81465d", + "comparison_retrieved_at": "2026-09-20T02:09:07.328339+00:00", + "status": "maintained snapshot is two commits ahead of release tag", + "changed_files": [ + "CITATION.cff", + "README.md", + "docs/blog.html", + "docs/index.html", + "paper/README.md" + ], + "limit": "GitHub compare reports documentation/citation differences only; keep exact source revisions distinct even though both declare version 0.2.1. No evaluator runtime equivalence test was executed." + }, + "subsequent_execution": { + "receipt": "blueprints/convergence-practice/arb-trace2code/receipt.json", + "source_commit": "b487f3866cc13dd971819cb902517a6a50282404", + "scope": "Separately recorded native CLI replay of the complete 101-case trace2code release; no installed retriever, agent repair or model acceptance. runtime_executed above refers only to this source-review wave." + } + }, + { + "repository": "UKGovernmentBEIS/inspect_ai", + "repository_key": "ukgovernmentbeis/inspect_ai", + "version": null, + "published_package_version": "0.3.266", + "source_commit": "ec4dfc6953784dc45b79de3147530c89868c6e26", + "retrieved_at": "2026-09-20T02:02:08.113262+00:00", + "release_url": null, + "published_at": null, + "source_depth": "pinned_readme_license_and_selected_supporting_source", + "license_at_pin": { + "spdx": "MIT", + "url": "https://github.com/UKGovernmentBEIS/inspect_ai/blob/ec4dfc6953784dc45b79de3147530c89868c6e26/LICENSE", + "sha256": "c593c2afc81388521eaeb03f0d1a4b01bc3e5147ed9cb16c1f129e5be353b6af" + }, + "catalog_coverage": { + "repository": "ukgovernmentbeis/inspect_ai", + "in_public_decision_index": true, + "public_references": [ + { + "kind": "catalog_card", + "path": "catalogs/us-equities/agents-operations.json", + "pointer": "/entries/30", + "id": "inspect-ai", + "decision": "default", + "evidence_level": "source_review" + } + ], + "public_star": false, + "in_maintained_registry": true, + "maintained_selection": "trial", + "maintained_pin": "inspect-ai0.3.266; source ec4dfc6953784dc45b79de3147530c89868c6e26" + }, + "discovery_provenance": [ + { + "kind": "public_owned_repository_catalog", + "source_url": "https://github.com/seathatflowsinourveins/native-agent-stack/blob/bf99d340f798bf72cc3104619edc482882eb423d/catalogs/us-equities/decision-index.json" + } + ], + "sources": [ + { + "path": "LICENSE", + "url": "https://github.com/UKGovernmentBEIS/inspect_ai/blob/ec4dfc6953784dc45b79de3147530c89868c6e26/LICENSE", + "blob_sha": "72fc87742ef8a944fab4f28fe6231696c62f2fa4", + "sha256": "c593c2afc81388521eaeb03f0d1a4b01bc3e5147ed9cb16c1f129e5be353b6af", + "bytes": 1081 + }, + { + "path": "README.md", + "url": "https://github.com/UKGovernmentBEIS/inspect_ai/blob/ec4dfc6953784dc45b79de3147530c89868c6e26/README.md", + "blob_sha": "1e58c142b2e05faf2e3dba9a487450c0845c4576", + "sha256": "a9a1e07b955d869137d8ad109f1c238268f361dcdfac30cfcc94983a51a99b17", + "bytes": 3099 + }, + { + "path": "pyproject.toml", + "url": "https://github.com/UKGovernmentBEIS/inspect_ai/blob/ec4dfc6953784dc45b79de3147530c89868c6e26/pyproject.toml", + "blob_sha": "ee90d00ac8edbeda0caf7ab56798e1b901dd47d5", + "sha256": "ce5713f4a354c3d2fd3b01f08e247a06b240a21eda481581b3cd270c867c1d2b", + "bytes": 5660 + }, + { + "path": "docs/eval-sets.qmd", + "url": "https://github.com/UKGovernmentBEIS/inspect_ai/blob/ec4dfc6953784dc45b79de3147530c89868c6e26/docs/eval-sets.qmd", + "blob_sha": "d09b2b9e05e5c4fc91e35990d7176f2ad1ac112c", + "sha256": "1f12dea4705f63125dd127700cdb10c8d11eb0aaf1f468ff3a5ccb92c505d951", + "bytes": 14758 + }, + { + "path": "docs/eval-logs.qmd", + "url": "https://github.com/UKGovernmentBEIS/inspect_ai/blob/ec4dfc6953784dc45b79de3147530c89868c6e26/docs/eval-logs.qmd", + "blob_sha": "c1a81f66e3ac5ba14728233c9f29ef0c811377c1", + "sha256": "09b38f97b74f36ef753db9fdfaeb2245a804347dcbc01d6525e69ff2dcc0a9df", + "bytes": 34291 + }, + { + "path": "src/inspect_ai/model/_providers/mockllm.py", + "url": "https://github.com/UKGovernmentBEIS/inspect_ai/blob/ec4dfc6953784dc45b79de3147530c89868c6e26/src/inspect_ai/model/_providers/mockllm.py", + "blob_sha": "7fb658b227093dbd1a1b735b91cc6a3bd8469ba6", + "sha256": "7b4038e930b8669389f51bc40b2353ba18ca561fc1c79ac6bf7d88f250dbdb2c", + "bytes": 5562 + } + ], + "runtime_executed": false, + "lane": "evaluation/recovery", + "decision": "investigate", + "why": "Provides reusable scoring, logs and resume/retry semantics for a small deterministic evaluation suite.", + "overlap": "Use it as an evaluation runner only; keep production workflow ownership and native client authentication unchanged.", + "compatibility": "Python >=3.10. Source reviewed at exact commit; PyPI latest 0.3.266 is separate package metadata, not a reproduced wheel/source equivalence check.", + "acceptance": "Run a tiny public synthetic suite through mockllm with deterministic scoring, max_tasks=1, bounded retry attempts and failure-log retention. Interrupt then resume; verify completed samples are reused, failure provenance remains and no provider traffic occurs. Only then consider an explicitly authorized provider lane.", + "limits": [ + "Default eval-set concurrency and ten retry attempts are broader than the proposed smoke test.", + "Eval logs can contain prompts/tool data; only sanitized summaries belong in public evidence.", + "Head source review is not installed-runtime qualification." + ], + "package_metadata": { + "url": "https://pypi.org/pypi/inspect_ai/0.3.266/json", + "retrieved_at": "2026-09-20T02:08:06.220557+00:00", + "version": "0.3.266", + "requires_python": ">=3.10", + "license": "MIT License", + "scope": "Package metadata only; artifact not downloaded or executed; source equivalence not independently reproduced." + }, + "review_level": "source_review", + "evidence_level": "pinned_source_review" + }, + { + "repository": "harbor-framework/harbor", + "repository_key": "harbor-framework/harbor", + "version": "v0.23.0", + "published_package_version": null, + "source_commit": "1e5c5c6db929a10a140d05e606882c671ae20729", + "retrieved_at": "2026-09-20T02:02:08.179873+00:00", + "release_url": "https://github.com/harbor-framework/harbor/releases/tag/v0.23.0", + "published_at": "2026-09-12T04:55:17Z", + "source_depth": "pinned_readme_license_and_selected_supporting_source", + "license_at_pin": { + "spdx": "Apache-2.0", + "url": "https://github.com/harbor-framework/harbor/blob/1e5c5c6db929a10a140d05e606882c671ae20729/LICENSE", + "sha256": "c71d239df91726fc519c6eb72d318ec65820627232b2f796219e87dcf35d0ab4" + }, + "catalog_coverage": { + "repository": "harbor-framework/harbor", + "in_public_decision_index": false, + "public_references": [], + "public_star": false, + "in_maintained_registry": false, + "maintained_selection": null, + "maintained_pin": null + }, + "discovery_provenance": [ + { + "kind": "primary_search", + "source_url": "https://github.com/harbor-framework/harbor/releases/tag/v0.23.0" + } + ], + "sources": [ + { + "path": "LICENSE", + "url": "https://github.com/harbor-framework/harbor/blob/1e5c5c6db929a10a140d05e606882c671ae20729/LICENSE", + "blob_sha": "261eeb9e9f8b2b4b0d119366dda99c6fd7d35c64", + "sha256": "c71d239df91726fc519c6eb72d318ec65820627232b2f796219e87dcf35d0ab4", + "bytes": 11357 + }, + { + "path": "README.md", + "url": "https://github.com/harbor-framework/harbor/blob/1e5c5c6db929a10a140d05e606882c671ae20729/README.md", + "blob_sha": "4ae345184f65e2c0ff8cce8e496a8fdcc439f9bc", + "sha256": "687e4f416b43e1ab28afb3cd48f556d7a9f4d664ac78ac1a2a07a7ab9b95b1b7", + "bytes": 2984 + }, + { + "path": "pyproject.toml", + "url": "https://github.com/harbor-framework/harbor/blob/1e5c5c6db929a10a140d05e606882c671ae20729/pyproject.toml", + "blob_sha": "f37c92e1b2de2dec7a042bad8db0f6483f9a76f3", + "sha256": "00290ec17fb5f4f44ebd1a251dd07fe4f1b6dabb9fed61497cd303a4679e104d", + "bytes": 5829 + }, + { + "path": "src/harbor/agents/oracle.py", + "url": "https://github.com/harbor-framework/harbor/blob/1e5c5c6db929a10a140d05e606882c671ae20729/src/harbor/agents/oracle.py", + "blob_sha": "70d7ef719198e00b89378d95cf468124f494200e", + "sha256": "9bebb320fd3909aea0fe7da174b4cb7c9b761591224d9db8010d73d1a002915c", + "bytes": 5867 + } + ], + "runtime_executed": false, + "lane": "evaluation/isolation", + "decision": "investigate", + "why": "Adds standardized agent task environments and verifiers when evaluating real terminal work, beyond file-retrieval scores.", + "overlap": "A genuinely new catalog identity, but overlaps existing native workers if used as another everyday orchestration runtime. Evaluate only as an optional benchmark harness.", + "compatibility": "Python >=3.12; local README example requires Docker. Provider examples use API keys; native subscription sign-in reuse was not verified.", + "acceptance": "First qualify one self-authored public task with the built-in oracle agent and deterministic verifier, concurrency=1 and explicit timeout; record container cleanup and preserved trial artifacts. This proves task packaging only. Defer model-agent benchmarking until a task requires it and credentials/budget are separately authorized.", + "limits": [ + "Container/oracle success is not agent or model success.", + "Do not infer equivalent credential access inside containers.", + "No Docker installation, benchmark environment or provider run occurred." + ], + "qualification_gate": "defer_pending_specific_benchmark", + "review_level": "source_review", + "evidence_level": "pinned_source_review" + }, + { + "repository": "ml-explore/mlx-lm", + "repository_key": "ml-explore/mlx-lm", + "version": "v0.31.3", + "published_package_version": null, + "source_commit": "ed1fca4cef15a824c5f1702c80f70b4cffc8e4dd", + "retrieved_at": "2026-09-20T02:02:08.096710+00:00", + "release_url": "https://github.com/ml-explore/mlx-lm/releases/tag/v0.31.3", + "published_at": "2026-04-22T07:43:57Z", + "source_depth": "pinned_readme_license_and_selected_supporting_source", + "license_at_pin": { + "spdx": "MIT", + "url": "https://github.com/ml-explore/mlx-lm/blob/ed1fca4cef15a824c5f1702c80f70b4cffc8e4dd/LICENSE", + "sha256": "ccfab7ccb2ea306f71531c8ca77bb55507606cd90768b1e32b8b52ab5b48cf01" + }, + "catalog_coverage": { + "repository": "ml-explore/mlx-lm", + "in_public_decision_index": true, + "public_references": [ + { + "kind": "research_supplement", + "path": "catalogs/us-equities/architecture/foundation.json", + "pointer": "/repositories/12", + "decision": "investigate", + "evidence_depth": "pinned_primary_source_review" + } + ], + "public_star": false, + "in_maintained_registry": true, + "maintained_selection": "discovery", + "maintained_pin": null + }, + "discovery_provenance": [ + { + "kind": "public_owned_repository_catalog", + "source_url": "https://github.com/seathatflowsinourveins/native-agent-stack/blob/bf99d340f798bf72cc3104619edc482882eb423d/catalogs/us-equities/decision-index.json" + } + ], + "sources": [ + { + "path": "LICENSE", + "url": "https://github.com/ml-explore/mlx-lm/blob/ed1fca4cef15a824c5f1702c80f70b4cffc8e4dd/LICENSE", + "blob_sha": "98ff47b9ef9d4ac9f1a4bde3db13dd27456c5ea1", + "sha256": "ccfab7ccb2ea306f71531c8ca77bb55507606cd90768b1e32b8b52ab5b48cf01", + "bytes": 1066 + }, + { + "path": "README.md", + "url": "https://github.com/ml-explore/mlx-lm/blob/ed1fca4cef15a824c5f1702c80f70b4cffc8e4dd/README.md", + "blob_sha": "ce71596b3cf0a879f1c7cbd9ab29632e08b1af16", + "sha256": "d7bfcf00a0a2e0a6d4a5da48eb523bbf0ad248b3f7c5511a20a72175e2958084", + "bytes": 8193 + }, + { + "path": "setup.py", + "url": "https://github.com/ml-explore/mlx-lm/blob/ed1fca4cef15a824c5f1702c80f70b4cffc8e4dd/setup.py", + "blob_sha": "dd769a706d7d7ea388a540aacb3599afd858f938", + "sha256": "d8049945cbe187386dab23122a2aa3c2a035aecc5a004033dd22f26064bca729", + "bytes": 2368 + } + ], + "runtime_executed": false, + "lane": "local-inference", + "decision": "investigate", + "why": "Apple-native generation/fine-tuning and prompt-cache controls may help a measured local workload.", + "overlap": "Existing Ollama covers local embeddings; MLX uses a separate model format/cache and is not a drop-in QMD GGUF backend.", + "compatibility": "Apple Silicon MLX; package setup declares Python >=3.8 but dependency wheels govern effective support. Wired-memory optimization requires macOS >=15.", + "acceptance": "Pick one extraction or bounded code-review workload with held-out answers. Compare locked model revisions/quantizations to the current route; measure accuracy, cold/warm time, tokens/sec and peak memory without changing system memory limits. Adopt only if a predefined quality/resource threshold is met.", + "limits": [ + "Runtime MIT license does not establish model-weight license.", + "Matching parameter count is not matching model/quantization.", + "No local model download or inference performed." + ], + "qualification_gate": "investigate_only_for_named_workload", + "review_level": "source_review", + "evidence_level": "pinned_source_review" + }, + { + "repository": "ggml-org/llama.cpp", + "repository_key": "ggml-org/llama.cpp", + "version": "v0.4.1", + "published_package_version": null, + "source_commit": "b29c606e28a01b1bc8c1351026a0fa6e616bf6c4", + "retrieved_at": "2026-09-20T02:02:08.108295+00:00", + "release_url": "https://github.com/ggml-org/llama.cpp/releases/tag/v0.4.1", + "published_at": "2026-09-14T18:27:29Z", + "source_depth": "pinned_readme_license_and_selected_supporting_source", + "license_at_pin": { + "spdx": "MIT", + "url": "https://github.com/ggml-org/llama.cpp/blob/b29c606e28a01b1bc8c1351026a0fa6e616bf6c4/LICENSE", + "sha256": "94f29bbed6a22c35b992c5c6ebf0e7c92f13b836b90f36f461c9cf2f0f1d010d" + }, + "catalog_coverage": { + "repository": "ggml-org/llama.cpp", + "in_public_decision_index": false, + "public_references": [], + "public_star": false, + "in_maintained_registry": true, + "maintained_selection": "trial", + "maintained_pin": "v0.4.1" + }, + "discovery_provenance": [ + { + "kind": "maintained_registry", + "source_url": "https://github.com/ggml-org/llama.cpp/releases/tag/v0.4.1" + } + ], + "sources": [ + { + "path": "LICENSE", + "url": "https://github.com/ggml-org/llama.cpp/blob/b29c606e28a01b1bc8c1351026a0fa6e616bf6c4/LICENSE", + "blob_sha": "e7dca554bcb802f98408383a864404e3aa4eacca", + "sha256": "94f29bbed6a22c35b992c5c6ebf0e7c92f13b836b90f36f461c9cf2f0f1d010d", + "bytes": 1078 + }, + { + "path": "README.md", + "url": "https://github.com/ggml-org/llama.cpp/blob/b29c606e28a01b1bc8c1351026a0fa6e616bf6c4/README.md", + "blob_sha": "aae3bcd35ad9c2ba3e914750a36e125fec0b5355", + "sha256": "e58520857253729851cc02313883c9d2a8b2ae86abe212aaf1bcbd7db8bfefb4", + "bytes": 7351 + }, + { + "path": "pyproject.toml", + "url": "https://github.com/ggml-org/llama.cpp/blob/b29c606e28a01b1bc8c1351026a0fa6e616bf6c4/pyproject.toml", + "blob_sha": "0383fbc5e6d409f14c13096cf666b1dd1f65d9b6", + "sha256": "03c52669f756ba7084659fd55467ec5f4990866d1c12f5a01833c4d56430df32", + "bytes": 1907 + } + ], + "runtime_executed": false, + "lane": "local-inference", + "decision": "investigate", + "why": "Direct GGUF/Metal execution and serving controls provide a transparent comparison lane for local generation.", + "overlap": "New to this public decision index but already a trial in the maintained registry; existing Ollama overlaps general local serving.", + "compatibility": "Apple Silicon is supported via NEON, Accelerate and Metal; stable v0.4.1 source reviewed. Python conversion scripts have their own runtime dependencies and are not the inference binary.", + "acceptance": "Use one exact GGUF artifact and prompt set, bounded context and concurrency=1. Compare quality, prompt processing/generation speed and resident memory with the existing route; verify cancellation releases the process/model. Do not add a persistent server until useful.", + "limits": [ + "Model artifact license/architecture must be checked separately.", + "Newer b-series prereleases are not this stable source pin.", + "No binary installation, server or model run performed." + ], + "qualification_gate": "investigate_only_for_named_workload", + "review_level": "source_review", + "evidence_level": "pinned_source_review" + }, + { + "repository": "restic/restic", + "repository_key": "restic/restic", + "version": "v0.19.1", + "published_package_version": null, + "source_commit": "6aa3a516ce654808a1f28f9fa21e9b7c8e6e90bf", + "retrieved_at": "2026-09-20T02:02:08.159288+00:00", + "release_url": "https://github.com/restic/restic/releases/tag/v0.19.1", + "published_at": "2026-07-05T08:13:33Z", + "source_depth": "pinned_readme_license_and_selected_supporting_source", + "license_at_pin": { + "spdx": "BSD-2-Clause", + "url": "https://github.com/restic/restic/blob/6aa3a516ce654808a1f28f9fa21e9b7c8e6e90bf/LICENSE", + "sha256": "6f08a01a9fab5b24e139a09f15cc24a73087c7bc09e3bacf099fdf2d767bf897" + }, + "catalog_coverage": { + "repository": "restic/restic", + "in_public_decision_index": true, + "public_references": [ + { + "kind": "catalog_card", + "path": "catalogs/us-equities/agents-operations.json", + "pointer": "/entries/41", + "id": "restic", + "decision": "default", + "evidence_level": "native_proven" + }, + { + "kind": "research_supplement", + "path": "catalogs/us-equities/convergence-program/hosting.json", + "pointer": "/candidates/9", + "decision": "retain", + "evidence_depth": "selected_primary_source_review" + }, + { + "kind": "public_star", + "path": "catalogs/us-equities/coverage.json", + "pointer": "/stars/245", + "disposition": "catalog_reviewed" + }, + { + "kind": "star_review", + "path": "catalogs/us-equities/star-audit.json", + "pointer": "/entries/175", + "decision": "targeted_candidate", + "review_level": "source_review", + "review_depth": "selected_primary_files" + }, + { + "kind": "component_record", + "path": "manifests/stack.json", + "pointer": "/components/43", + "id": "restic" + } + ], + "public_star": true, + "in_maintained_registry": true, + "maintained_selection": "selected", + "maintained_pin": "0.19.1" + }, + "discovery_provenance": [ + { + "kind": "public_owned_repository_catalog", + "source_url": "https://github.com/seathatflowsinourveins/native-agent-stack/blob/bf99d340f798bf72cc3104619edc482882eb423d/catalogs/us-equities/decision-index.json" + }, + { + "kind": "public_star", + "source_url": "https://api.github.com/users/seathatflowsinourveins/starred?per_page=100", + "retrieved_at": "2026-09-20T02:00:50.538618+00:00" + }, + { + "kind": "awesome_list", + "source_url": "https://github.com/awesome-foss/awesome-sysadmin/blob/a0f64294fa6c672767f4dffc978023194bf99550/README.md#L111" + } + ], + "sources": [ + { + "path": "LICENSE", + "url": "https://github.com/restic/restic/blob/6aa3a516ce654808a1f28f9fa21e9b7c8e6e90bf/LICENSE", + "blob_sha": "29b2ee55a92ff2c5516d7b916224c85fa7e9139a", + "sha256": "6f08a01a9fab5b24e139a09f15cc24a73087c7bc09e3bacf099fdf2d767bf897", + "bytes": 1345 + }, + { + "path": "README.md", + "url": "https://github.com/restic/restic/blob/6aa3a516ce654808a1f28f9fa21e9b7c8e6e90bf/README.md", + "blob_sha": "ef12f3e1b2a2992c1b2d5d183c57c828891f4f79", + "sha256": "8180046c28e55384c37588607b64a160b3e359a044d6ac59fddb94d659eab75d", + "bytes": 5706 + }, + { + "path": "go.mod", + "url": "https://github.com/restic/restic/blob/6aa3a516ce654808a1f28f9fa21e9b7c8e6e90bf/go.mod", + "blob_sha": "00e1171de5d2a17f21d2d13f9024ef2956e6afaa", + "sha256": "d166913ed4897cf69069ef5be090b2d79a857fc28328bb950e12655b5a54f176", + "bytes": 4894 + } + ], + "runtime_executed": false, + "lane": "recovery", + "decision": "retain", + "why": "The unresolved capability is independent recovery, not another backup engine.", + "overlap": "Already selected and source-host-proven for narrower restore scopes; avoid another orchestrator before off-host restore is demonstrated.", + "compatibility": "Upstream supports Linux, macOS and Windows; credentials and repository keys stay in native/private stores.", + "acceptance": "Using an explicitly selected independent destination and separately recoverable key, back up a non-sensitive fixture, restore into a clean target, compare content hashes and permissions, then simulate loss of original working files. Record recovery time and missing-key failure without publishing secrets. Never delete real originals for the drill.", + "limits": [ + "Same-host backup/restore does not prove disaster recovery.", + "README makes lost repository passwords unrecoverable; key recovery is a separate acceptance dependency.", + "No backup destination, account or service changed." + ], + "qualification_gate": "retain_and_qualify_independent_restore", + "review_level": "source_review", + "evidence_level": "pinned_source_review" + }, + { + "repository": "tobi/qmd", + "repository_key": "tobi/qmd", + "version": "v2.8.3", + "published_package_version": null, + "source_commit": "facd35e01359e59d938bc9418e93fb9318addee3", + "retrieved_at": "2026-09-20T02:02:10.030799+00:00", + "release_url": "https://github.com/tobi/qmd/releases/tag/v2.8.3", + "published_at": "2026-08-16T22:30:34Z", + "source_depth": "pinned_readme_license_and_selected_supporting_source", + "license_at_pin": { + "spdx": "MIT", + "url": "https://github.com/tobi/qmd/blob/facd35e01359e59d938bc9418e93fb9318addee3/LICENSE", + "sha256": "24c446836e7e2cdea13a914de10c88c553e6297366061ba24013b8ce73c8f7fa" + }, + "catalog_coverage": { + "repository": "tobi/qmd", + "in_public_decision_index": true, + "public_references": [ + { + "kind": "research_supplement", + "path": "catalogs/us-equities/architecture/foundation.json", + "pointer": "/repositories/2", + "decision": "retain", + "evidence_depth": "pinned_primary_source_review" + }, + { + "kind": "research_supplement", + "path": "catalogs/us-equities/convergence-program/foundation.json", + "pointer": "/repositories/3", + "decision": "retain", + "evidence_level": "pinned_primary_source_reuse" + }, + { + "kind": "catalog_card", + "path": "catalogs/us-equities/foundation-memory.json", + "pointer": "/entries/3", + "id": "foundation-qmd", + "decision": "default", + "evidence_level": "native_proven" + }, + { + "kind": "legacy_candidate", + "path": "manifests/candidates.json", + "pointer": "/candidates/7", + "decision": "First evaluate semantic mode of existing install" + }, + { + "kind": "component_record", + "path": "manifests/stack.json", + "pointer": "/components/25", + "id": "qmd" + } + ], + "public_star": false, + "in_maintained_registry": true, + "maintained_selection": "selected", + "maintained_pin": "2.8.3" + }, + "discovery_provenance": [ + { + "kind": "public_owned_repository_catalog", + "source_url": "https://github.com/seathatflowsinourveins/native-agent-stack/blob/bf99d340f798bf72cc3104619edc482882eb423d/catalogs/us-equities/decision-index.json" + } + ], + "sources": [ + { + "path": "LICENSE", + "url": "https://github.com/tobi/qmd/blob/facd35e01359e59d938bc9418e93fb9318addee3/LICENSE", + "blob_sha": "81652d05896680c9c78e5f4a3bf69cf719f5818c", + "sha256": "24c446836e7e2cdea13a914de10c88c553e6297366061ba24013b8ce73c8f7fa", + "bytes": 1072 + }, + { + "path": "README.md", + "url": "https://github.com/tobi/qmd/blob/facd35e01359e59d938bc9418e93fb9318addee3/README.md", + "blob_sha": "2c697f3ee9cdb7f53aea1b41342563a88887f7d4", + "sha256": "24c049f59408d3d021fbee70d4cd5dd744cdb51f5caea956b417f4dfe7e26880", + "bytes": 50056 + }, + { + "path": "package.json", + "url": "https://github.com/tobi/qmd/blob/facd35e01359e59d938bc9418e93fb9318addee3/package.json", + "blob_sha": "1c6082ab9783d0d76042f102d517b04b30e2be88", + "sha256": "14462e2764e9d140afbeb3c924e6dd81c0448177e97785822e763d901d70f20b", + "bytes": 3415 + } + ], + "runtime_executed": false, + "lane": "context/retrieval", + "decision": "retain", + "why": "Existing selected document search already has lexical/vector/hybrid/full benchmark modes; measure the unqualified lane before adding another document engine.", + "overlap": "Keep document collections explicit; SocratiCode remains code-context retrieval and ai-memory remains explicit durable decisions.", + "compatibility": "Node >=22. Semantic/reranking/query expansion use specific GGUF contracts via node-llama-cpp and may download models on first use; no model switch is implied.", + "acceptance": "Confirm the named collection is populated, freeze unseen lexical/semantic/alias/no-answer queries and expected files, then compare BM25, vector, hybrid and full pipelines. Measure relevance and latency with exact model artifacts. Preserve the old lexical baseline and rollback index.", + "limits": [ + "Upstream bench can report all zeros without warning for an absent collection, so setup must be checked.", + "Upstream example scores do not establish the user corpus result.", + "No index refresh, broad directory scan or model download occurred." + ], + "qualification_gate": "retain_and_measure_existing_semantic_lane", + "review_level": "source_review", + "evidence_level": "pinned_source_review" + } + ], + "summary": { + "candidates_reviewed": 7, + "already_in_public_index": 5, + "new_to_public_index": [ + "ggml-org/llama.cpp", + "harbor-framework/harbor" + ], + "genuinely_new_to_maintained_registry": [ + "harbor-framework/harbor" + ], + "proposed_adoptions": 0 + }, + "limits": [ + "GET-only public source acquisition; no installation, configuration, model call, benchmark execution, publication or repository mutation.", + "Public inventory counts reflect endpoint visibility at retrieval time; private repository identities and account stores were not inspected.", + "Awesome lists supplied specific discovery links only; no entire-list audit, popularity ranking or copied quality claim.", + "Exact source and license review does not prove dependency resolution, binary equivalence, hardware performance or account entitlement.", + "Historical source-host receipts remain separate from current target-host execution.", + "The maintained registry comparison records only public candidate identities, status and pins; no private repository identity or location is included." + ] +} diff --git a/catalogs/convergence-practice/source-review.md b/catalogs/convergence-practice/source-review.md new file mode 100644 index 000000000..5ab2d58f0 --- /dev/null +++ b/catalogs/convergence-practice/source-review.md @@ -0,0 +1,27 @@ +# Bounded public convergence review + +Observed 2026-09-20T02:07:12.602467+00:00. Public base `bf99d340f798bf72cc3104619edc482882eb423d`. No installations, model calls or runtime acceptance. + +Public inventory: **1 owned repository, 342 starred repositories**, all stars already present in the public decision index. Three pinned starred awesome lists were inspected for relevant entries; discovery links were checked separately from primary sources. Star counts were not used to rank quality. + +| Candidate / source pin | License | Existing coverage | Recommended next acceptance | +|---|---|---|---| +| [eyuansu62/agent-retrieval-bench v0.2.1](https://github.com/eyuansu62/agent-retrieval-bench/releases/tag/v0.2.1) | MIT | Already public | Freeze a stratified public subset and hashes before runs. Compare lexical, existing semantic and selective-abstention lanes under the same all_files candidates and 4k/8k token budgets. Report Recall@k, MRR, BCY, no-gold errors, time and actual provider usage separately. Add independent edit/test tasks before claiming patch improvement. | +| [UKGovernmentBEIS/inspect_ai source ec4dfc695378](https://github.com/UKGovernmentBEIS/inspect_ai/tree/ec4dfc6953784dc45b79de3147530c89868c6e26) | MIT | Already public | Run a tiny public synthetic suite through mockllm with deterministic scoring, max_tasks=1, bounded retry attempts and failure-log retention. Interrupt then resume; verify completed samples are reused, failure provenance remains and no provider traffic occurs. Only then consider an explicitly authorized provider lane. | +| [harbor-framework/harbor v0.23.0](https://github.com/harbor-framework/harbor/releases/tag/v0.23.0) | Apache-2.0 | Genuinely new candidate | First qualify one self-authored public task with the built-in oracle agent and deterministic verifier, concurrency=1 and explicit timeout; record container cleanup and preserved trial artifacts. This proves task packaging only. Defer model-agent benchmarking until a task requires it and credentials/budget are separately authorized. | +| [ml-explore/mlx-lm v0.31.3](https://github.com/ml-explore/mlx-lm/releases/tag/v0.31.3) | MIT | Already public | Pick one extraction or bounded code-review workload with held-out answers. Compare locked model revisions/quantizations to the current route; measure accuracy, cold/warm time, tokens/sec and peak memory without changing system memory limits. Adopt only if a predefined quality/resource threshold is met. | +| [ggml-org/llama.cpp v0.4.1](https://github.com/ggml-org/llama.cpp/releases/tag/v0.4.1) | MIT | Existing maintained trial; new public identity | Use one exact GGUF artifact and prompt set, bounded context and concurrency=1. Compare quality, prompt processing/generation speed and resident memory with the existing route; verify cancellation releases the process/model. Do not add a persistent server until useful. | +| [restic/restic v0.19.1](https://github.com/restic/restic/releases/tag/v0.19.1) | BSD-2-Clause | Already public | Using an explicitly selected independent destination and separately recoverable key, back up a non-sensitive fixture, restore into a clean target, compare content hashes and permissions, then simulate loss of original working files. Record recovery time and missing-key failure without publishing secrets. Never delete real originals for the drill. | +| [tobi/qmd v2.8.3](https://github.com/tobi/qmd/releases/tag/v2.8.3) | MIT | Already public | Confirm the named collection is populated, freeze unseen lexical/semantic/alias/no-answer queries and expected files, then compare BM25, vector, hybrid and full pipelines. Measure relevance and latency with exact model artifacts. Preserve the old lexical baseline and rollback index. | + +The strongest immediate moves are **ARB plus a separate edit-success test**, **a small resumable Inspect suite**, **independent restic recovery**, and **QMD semantic evaluation**. These turn existing components into measured capabilities. Local MLX/llama comparisons should follow a named workload and resource ceiling. Harbor earns a conditional benchmark-harness review, not another default runtime. + +Inspection found useful limits: ARB file retrieval is not patch success; its corpus keeps upstream licenses. Inspect eval-set defaults can expand retries/concurrency and clean failed logs, so the proposed smoke test must explicitly bound these and retain failure evidence. QMD bench can return zeros for a missing collection. Harbor examples use Docker and API-key routes, so native account reuse is unproven. + +The awesome-list links were concrete: restic/Borg/Backrest from awesome-sysadmin; Serena/memo from awesome-mcp-servers; SWE-agent from awesome-ai-agents. The latter alternatives remain discovery-only because they duplicate an existing role or have no demonstrated gap. Exact list commits, licenses, line links and source hashes are recorded in source-review.json. + +**No adoption is proposed solely from this wave.** Five candidates already appear in the public index; llama.cpp already exists in the maintained registry; Harbor is the only genuinely new candidate. Neither fresh metadata nor these source reviews establishes a SOTA winner or new host acceptance. + +ARB source distinction: this wave inspected release tag `v0.2.1` at `b487f3866cc13dd971819cb902517a6a50282404`. The maintained source snapshot `07014c986f3deadb1548c62b32c0ffbe6a81465d` is two commits ahead; [the exact comparison](https://github.com/eyuansu62/agent-retrieval-bench/compare/b487f3866cc13dd971819cb902517a6a50282404...07014c986f3deadb1548c62b32c0ffbe6a81465d) changes documentation/citation files only. Preserve both source pins in receipts instead of collapsing them into the shared declared version `0.2.1`. + +A subsequent, separately scoped [native ARB replay](../../blueprints/convergence-practice/arb-trace2code/README.md) executed both baselines at the exact release commit. The source-review status above remains historical to this review phase; the execution receipt records the measured results and limits. diff --git a/catalogs/us-equities/README.md b/catalogs/us-equities/README.md index d82b2126d..967928598 100644 --- a/catalogs/us-equities/README.md +++ b/catalogs/us-equities/README.md @@ -2,7 +2,7 @@ **Dated decision catalog: September 20, 2026.** The north star is native, token-efficient research → reproducible backtesting → Alpaca paper automation. This is an examined selection across layers, not a universal final SOTA ranking or a claim that every listed framework runs together. -The catalog has **152 baseline repository decision cards covering 147 unique GitHub repositories**, and **20 model entries**. The [combined repository index](repository-index.md) now contains **502 repository identities**, including all 342 public stars and 160 beyond that snapshot. Its [typed decision union](decision-index.json) validates baseline cards, component/candidate records and explicitly registered research supplements together; 1,030 source pointers preserve their different evidence depths. The historical 453-row index remains dated reference material. +The catalog has **152 baseline repository decision cards covering 147 unique GitHub repositories**, and **20 model entries**. The [combined repository index](repository-index.md) now contains **504 repository identities**, including all 342 public stars and 162 beyond that snapshot. Its [typed decision union](decision-index.json) validates baseline cards, component/candidate records and explicitly registered research supplements together; 1,037 source pointers preserve their different evidence depths. The historical 453-row index remains dated reference material. The newest [security-identity review](security-identity-review.md) examines Alpaca, Zipline, Qlib, NautilusTrader, LEAN and WRDS. It separates engine identity/lifetime diff --git a/catalogs/us-equities/decision-index.json b/catalogs/us-equities/decision-index.json index 420849457..9e62e6710 100644 --- a/catalogs/us-equities/decision-index.json +++ b/catalogs/us-equities/decision-index.json @@ -8,6 +8,12 @@ "repository_field": "repository", "kind": "research_supplement" }, + { + "path": "catalogs/convergence-practice/source-review.json", + "collection": "/candidates", + "repository_field": "repository", + "kind": "research_supplement" + }, { "path": "catalogs/us-equities/agents-operations.json", "collection": "/entries", @@ -143,16 +149,16 @@ "zep-ai/graphiti": "getzep/graphiti" }, "counts": { - "repositories": 502, + "repositories": 504, "public_star_repositories": 342, - "beyond_public_stars": 160, - "references": 1030, + "beyond_public_stars": 162, + "references": 1037, "repositories_by_record_type": { "catalog_card": 147, "component_record": 52, "legacy_candidate": 19, "public_star": 342, - "research_supplement": 96, + "research_supplement": 100, "star_review": 342 } }, @@ -4025,9 +4031,18 @@ "aliases": [], "public_star": false, "record_types": [ - "catalog_card" + "catalog_card", + "research_supplement" ], "references": [ + { + "kind": "research_supplement", + "path": "catalogs/convergence-practice/source-review.json", + "pointer": "/candidates/0", + "decision": "investigate", + "review_level": "source_review", + "evidence_level": "pinned_source_review" + }, { "kind": "catalog_card", "path": "catalogs/us-equities/foundation-memory.json", @@ -4521,6 +4536,24 @@ } ] }, + { + "repository": "https://github.com/ggml-org/llama.cpp", + "aliases": [], + "public_star": false, + "record_types": [ + "research_supplement" + ], + "references": [ + { + "kind": "research_supplement", + "path": "catalogs/convergence-practice/source-review.json", + "pointer": "/candidates/4", + "decision": "investigate", + "review_level": "source_review", + "evidence_level": "pinned_source_review" + } + ] + }, { "repository": "https://github.com/giancarloerra/socraticode", "aliases": [], @@ -4937,6 +4970,24 @@ } ] }, + { + "repository": "https://github.com/harbor-framework/harbor", + "aliases": [], + "public_star": false, + "record_types": [ + "research_supplement" + ], + "references": [ + { + "kind": "research_supplement", + "path": "catalogs/convergence-practice/source-review.json", + "pointer": "/candidates/2", + "decision": "investigate", + "review_level": "source_review", + "evidence_level": "pinned_source_review" + } + ] + }, { "repository": "https://github.com/headroomlabs-ai/headroom", "aliases": [ @@ -7034,6 +7085,14 @@ "research_supplement" ], "references": [ + { + "kind": "research_supplement", + "path": "catalogs/convergence-practice/source-review.json", + "pointer": "/candidates/3", + "decision": "investigate", + "review_level": "source_review", + "evidence_level": "pinned_source_review" + }, { "kind": "research_supplement", "path": "catalogs/us-equities/architecture/foundation.json", @@ -9609,6 +9668,14 @@ "star_review" ], "references": [ + { + "kind": "research_supplement", + "path": "catalogs/convergence-practice/source-review.json", + "pointer": "/candidates/5", + "decision": "retain", + "review_level": "source_review", + "evidence_level": "pinned_source_review" + }, { "kind": "catalog_card", "path": "catalogs/us-equities/agents-operations.json", @@ -11123,6 +11190,14 @@ "research_supplement" ], "references": [ + { + "kind": "research_supplement", + "path": "catalogs/convergence-practice/source-review.json", + "pointer": "/candidates/6", + "decision": "retain", + "review_level": "source_review", + "evidence_level": "pinned_source_review" + }, { "kind": "research_supplement", "path": "catalogs/us-equities/architecture/foundation.json", @@ -11584,9 +11659,18 @@ "aliases": [], "public_star": false, "record_types": [ - "catalog_card" + "catalog_card", + "research_supplement" ], "references": [ + { + "kind": "research_supplement", + "path": "catalogs/convergence-practice/source-review.json", + "pointer": "/candidates/1", + "decision": "investigate", + "review_level": "source_review", + "evidence_level": "pinned_source_review" + }, { "kind": "catalog_card", "path": "catalogs/us-equities/agents-operations.json", diff --git a/catalogs/us-equities/manifest.json b/catalogs/us-equities/manifest.json index 96d33e6da..196874af1 100644 --- a/catalogs/us-equities/manifest.json +++ b/catalogs/us-equities/manifest.json @@ -45,7 +45,7 @@ "harness_contract": "blueprints/us-equities/harness-contract.json", "broader_reference_union": 171, "star_audit_file": "catalogs/us-equities/star-audit.json", - "grand_index_repositories": 502, - "grand_index_beyond_stars": 160, + "grand_index_repositories": 504, + "grand_index_beyond_stars": 162, "decision_index_file": "catalogs/us-equities/decision-index.json" } diff --git a/manifests/evidence.json b/manifests/evidence.json index 20c95ff32..26fd34e6d 100644 --- a/manifests/evidence.json +++ b/manifests/evidence.json @@ -898,8 +898,8 @@ }, { "path": "README.md", - "sha256": "4e83851f41569ab576ab1df1022d472d5d946ffa54321f1be5a62779f707f0f6", - "bytes": 12982 + "sha256": "bc5427a7054fcddf04b541817ba7d75e2a00b879a2a11f70dcd698300b1df866", + "bytes": 13428 }, { "path": "adoption/README.md", @@ -986,6 +986,91 @@ "sha256": "22c6b2a8b39bde0ff2acdd300d4f17df9b3be71a08821d292d7066df47669af1", "bytes": 8234 }, + { + "path": "blueprints/convergence-practice/README.md", + "sha256": "c2fe9c4d70e5df43f5ea537ab9355c4afe2c3cafbad21a5dd40b46c1ac2fe429", + "bytes": 9616 + }, + { + "path": "blueprints/convergence-practice/arb-trace2code/README.md", + "sha256": "214f2f09a151792c6e83b6705eb1973d5efeb3c3b16be563ff7cd4cca0342e16", + "bytes": 7846 + }, + { + "path": "blueprints/convergence-practice/arb-trace2code/input-files.json", + "sha256": "2203d0f03d0ccb5e863e61df5c47d09023204fbbbf8cf51b1699a82e76cc1e25", + "bytes": 22815 + }, + { + "path": "blueprints/convergence-practice/arb-trace2code/metrics.json", + "sha256": "e2329e35dcd3ebe9fb831814264aa71bbd9edfc4172a487c79cb38971c03197a", + "bytes": 4133 + }, + { + "path": "blueprints/convergence-practice/arb-trace2code/paired-samples.json", + "sha256": "b44b3547e8db13f5acfc48ba828153f4833bd04e6afbbd9ce38e11c67dcd8286", + "bytes": 40461 + }, + { + "path": "blueprints/convergence-practice/arb-trace2code/receipt.json", + "sha256": "78a11243a78b3829e2bb6359a356a1b412a0d2211119b1b83b37d992af477855", + "bytes": 9334 + }, + { + "path": "blueprints/convergence-practice/arb-trace2code/source-files.json", + "sha256": "d87ac16f0e85d6ebdcba5adb519d3bf556253e2674c2a972258df4efa75657d2", + "bytes": 10140 + }, + { + "path": "blueprints/convergence-practice/arb-trace2code/upstream-bm25-summary.json", + "sha256": "1b2451800f0b81060486659a71dc40b447097c903299dbf9ea841019128789e9", + "bytes": 3685 + }, + { + "path": "blueprints/convergence-practice/arb-trace2code/upstream-lexical-summary.json", + "sha256": "6fe6c2746c3bc0a855ee452b2de59707730d199b6f3128c1f7dee4f0b928132b", + "bytes": 3684 + }, + { + "path": "blueprints/convergence-practice/arb-trace2code/upstream-validation.json", + "sha256": "2fdff3fa025e8d82e77ee23756d490a4579cd9d5f5796bf57ced2306dcb2d922", + "bytes": 129 + }, + { + "path": "blueprints/convergence-practice/local-fixture/corpus-manifest.json", + "sha256": "7bb986160b332e205208b6767b95b6f019423ffed1f50843534ff7e4272fdded", + "bytes": 4436 + }, + { + "path": "blueprints/convergence-practice/local-fixture/evaluation-receipt.json", + "sha256": "3887a25ad35c3b36af6e34ad2307a043fd1918e91a14f890e0e6cc230c6be6ba", + "bytes": 5237 + }, + { + "path": "blueprints/convergence-practice/local-fixture/evaluation.md", + "sha256": "695e61ebe524d0559280a7f3c9247c6e613355215387678ce76fd88411d76b9a", + "bytes": 10969 + }, + { + "path": "blueprints/convergence-practice/local-fixture/fixture.json", + "sha256": "aea5e9616fb19eb1d2fd58ecfadf88ff087d7d2e79b4a70bc9f88207af9d0402", + "bytes": 37835 + }, + { + "path": "blueprints/convergence-practice/local-fixture/freeze-receipt.json", + "sha256": "20033998bcdbf5879e7e6e062f02e7192344e85e7e8700e9b22fd53651b71bc7", + "bytes": 669 + }, + { + "path": "blueprints/convergence-practice/local-fixture/rankings.json", + "sha256": "35fc05c3483fa5f4a674ee190c543748c1f72d589d9b8f0ed83c2c667f8cd36e", + "bytes": 55347 + }, + { + "path": "blueprints/convergence-practice/protocol.json", + "sha256": "7176889bb4f8e3839c5230918c1e23ff89bdd3778a8f333db2ea881731a08e45", + "bytes": 2595 + }, { "path": "blueprints/us-equities/README.md", "sha256": "268f4d9e91d250177ed27324ea09e84149ece10a575c4102522b70f7bb26045d", @@ -2071,9 +2156,29 @@ "sha256": "743c1f63d87ea9c4186b1d569ff02fc3a9dd32b3a7283048201ffee32b481cf5", "bytes": 1309 }, + { + "path": "catalogs/convergence-practice/public-owned.json", + "sha256": "541528d9cf379a9bb0b7793d3984c62412d7753b4cb614df15d4d9d798a48f5f", + "bytes": 810 + }, + { + "path": "catalogs/convergence-practice/public-starred.json", + "sha256": "905362c3922a05148d4f715ce623ed426ec2a30139bb349ff50facd581b8014a", + "bytes": 227345 + }, + { + "path": "catalogs/convergence-practice/source-review.json", + "sha256": "e163cf9c85a83440bca0c0eec4a857712e1fc3dcbe2920289c1a09b73f06762c", + "bytes": 41480 + }, + { + "path": "catalogs/convergence-practice/source-review.md", + "sha256": "c0e915563472bdce42455978fde17276f482cb088ea77e852312720dd845c909", + "bytes": 6039 + }, { "path": "catalogs/us-equities/README.md", - "sha256": "89ba8cbaba3dc3df9de40ad815c6e4b9f6aa6220c0324415ce81e677ce3e718d", + "sha256": "483ad2424434b4346893ec5d6978ae11d0aac7f2e96156a886b6cf125b134b85", "bytes": 12611 }, { @@ -2218,8 +2323,8 @@ }, { "path": "catalogs/us-equities/decision-index.json", - "sha256": "21733b37d9841c1cb71ed7592a52b6b3071ed68a537260ca8fd7fc6c268c15d7", - "bytes": 376207 + "sha256": "cdfae5f79f31a0c79baf3ea09d7f4b47b16a9d509a2d5355bba8e22e19e710b3", + "bytes": 378992 }, { "path": "catalogs/us-equities/engines-strategies.json", @@ -2253,7 +2358,7 @@ }, { "path": "catalogs/us-equities/manifest.json", - "sha256": "c24776afe39b02698dc3adb9de0fdb471994cb99546aa8a85897d3a6937d7067", + "sha256": "80bda6a63665bab4a8fa6ea7383d1d9c89f87de2987e9f5c43605e514ea34681", "bytes": 2376 }, {