Repository navigation
Follow-up: finish Agent notification failure diagnostics - #12562
Conversation
Bugbot is paused — on-demand spend limit reachedBugbot uses usage-based billing for this team and has hit its on-demand spend limit. A team admin can raise the spend limit in the Cursor dashboard, or wait for the next billing cycle to continue. |
|
All contributors have signed the CLA ✍️ ✅ |
|
No actionable comments were generated in the recent review. 🎉 ℹ️ Recent review info⚙️ Run configurationConfiguration used: Path: .coderabbit.yaml Review profile: ASSERTIVE Plan: Advanced Run ID: 📒 Files selected for processing (3)
Included review availability: Your plan provides up to 10 included reviews per hour; 2 remain after this review. 📝 WalkthroughWalkthroughThe app-host classifier now detects numeric and named signal failures. Tests cover diagnosis edge cases, selected-suite outcomes, and log retention. The notification test workflow also runs when the classifier changes. ChangesApp-host failure classification
Priority: ⬇️ Low Estimated code review effort: 3 (Moderate) | ~20 minutes Change: Bug fix Merge Risk: ⚪ Minimal · up to The updated classifier recognizes the supported contextual signal crash output, with no remaining actionable merge risk identified. 🚥 Pre-merge checks | ✅ 24 | ❌ 1❌ Failed checks (1 warning)
✅ Passed checks (24 passed)
Full details: Docstring CoverageExplanation Docstring coverage is 10.00% which is insufficient. The required threshold is 80.00%. Docstring coverage is scoped to functions touched by this diff. Analyzed 10 functions across 2 files. (1 skipped: 1 unsupported.)
✨ Finishing Touches 💡 1📝 Generate docstrings 💡
🧪 Generate unit tests (beta)
Thanks for using CodeRabbit! It's free for OSS, and your support helps us grow. If you like it, consider giving us a shout-out. Comment |
Bugbot is paused — on-demand spend limit reachedBugbot uses usage-based billing for this team and has hit its on-demand spend limit. A team admin can raise the spend limit in the Cursor dashboard, or wait for the next billing cycle to continue. |
4638e5b Merge pull request manaflow-ai#12570 from manaflow-ai/issue-7272-undo-stack-crash 61eedef fix: isolate web undo targets before app menu routing c83dc7f Merge pull request manaflow-ai#12562 from manaflow-ai/issue-12532-agent-notification-flaky bfe1d3f Merge pull request manaflow-ai#12571 from manaflow-ai/issue-12547-nightly-provider-duplicates-guard f0d1635 chore: remove Cloud provider Release compile guard 509e806 Merge pull request manaflow-ai#12569 from manaflow-ai/issue-12567-cloud-machine-connectivity a5410da diagnostics(cloud): correlate machine terminal attachment state 7f6de01 test: reproduce application and markdown undo lifetime crashes 720f24a fix: close contextual signal matcher 50e3416 test: preserve numeric crash signal diagnosis d21502d test: exercise selected semantic suite reporting 3a0aff2 fix: retain contextual signal crash markers 16b9e75 test: preserve contextual signal crash diagnosis 557bb90 fix: trigger semantic workflow for classifier changes 46f0ba7 test: cover non-crash signal text 3824276 fix: avoid matching build signatures as signals 7079066 test: ignore build signatures in app-host causes # Conflicts: # .github/workflows/agent-notification-tests.yml
Follow-up to #12546
PR #12546 was merged before its reusable app-host check completed. This follow-up keeps #12532 open and carries the remaining fixes made afterward:
Agent notification semanticswhenscripts/ci/classify-app-host-test-output.pychanges, so classifier fixes are exercised by the exact workflow.*** Signal <number>and contextualSIGABRT/SIGSEGVcrash markers while ignoring incidental build text such asBuild description signature.Evidence
8f6fc5c8f9): attempt 1 app-host job 103661975255 failed after 76/5/4 tests with the pre-Fix Agent notification semantics CI reliability #12455 Claude fixture reportingstatus=0,timedOut=true; attempt 2 app-host job 103664897630 passed all five selected suites (76/5/4/9/3) on the same SHA and Xcode 26.3/macOS 15 arm64 toolchain. This proves the oldwaitUntilExit()observer race.720f24ade3), app-host job 103824957050:The test runner timed out while preparing to run tests; diagnosispre-test app-host failure,executed_tests=0, no semantic assertions.720f24ade3), app-host job 103829129669: the same pre-test app-host timeout on a different WarpBuild arm64 machine; diagnosispre-test app-host failure,executed_tests=0, no semantic assertions.The current repeated failures are hosted runner/testmanagerd setup defects, not semantic assertion failures. The workflow now preserves and reports that distinction and remains red when assertions never execute. The follow-up changes no notification product semantics and does not remove or weaken Claude/Codex semantic coverage.
Related to #12532
The test runner timed out while preparing to run testsfailure after ~25 minutes, withexecuted_tests=0and no semantic assertions. Three attempts on the same final SHA now agree on the pre-test app-host category.Handoff: the remaining red check belongs to the hosted XCTest/testmanagerd runner path used by the reusable workflow. Product notification semantics and package tests are green; no further notification code change is justified by these zero-test failures.