GitHub gh-aw: an agent’s failure-investigation ticket is not resolved by closing it

#maintenance #review #closure

GitHub gh-aw: an agent’s failure-investigation ticket is not resolved by closing it

Original artifacts: gh-aw issue #49022, opened July 30, 2026 by github-actions and marked generated by a six-hour Failure Investigator agent, plus its August 6 six-hour investigator report #50734. This is internal agent-workflow maintenance in a public repo, not a customer support ticket or proof an external requester was served. The investigator’s text and linked run evidence are primary operational records but not an independent root-cause audit.

Trace. The original report cites two PR-review-workflow runs in which a background Copilot review subagent stalled (parent timed out) or failed model allocation on repeated retries; it asks for bounded subagent waiting, nonfatal fallback and no fourfold whole-session replay. Subsequent generated updates link recurrences in other workflows and enlarge the scope from one reviewer to background subagents across the fleet. The ticket was closed ‘not planned’ August 6 at 03:45:57 UTC; the next six-hour report says it reopened the same ticket at 07:41 after six further runs across three workflows. Verification boundary: the report log-fetched one of the six and explicitly labeled five as presumed same signature, so six is not six individually confirmed cause matches. Another agent-generated update says it was again closed ‘not planned’ at 09:31 on August 6 and reopened after three more runs by August 7; the public issue currently shows ‘Closed as not planned.’ This record documents contest over state, not a demonstrated merged fix or terminal success.

Why it matters. The monitoring agent could recognize recurrence, update its scope, and assert an issue should reopen when failures persisted. But an issue status and an agent assertion were not the same as repair; repeated administrative close/reopen cycles, triage costs and running failures are part of the economics of delegated maintenance. It also shows that a missing or failed secondary reviewer needs a visible partial-degradation policy, not silently successful green checks. Do not count these run failures as customer regressions or blame a person/agent for the closures without an event timeline attributing the actions. See requester-confirmation test, follow-up repair, review authority.