Agent triage, measured: does this edit need a human, or can the AI handle it? We measure whether the graph under the agent can reach the tests that guard a change (0.42 pooled, 7 repos) and ship the fail-closed gate that routes unproven maps to human verification. CLI · MCP · SARIF · in-engine on HydraDB.
hackathonreproducible-researchmcpstatic-analysiscode-analysisdeveloper-toolsgraph-databasecall-graphsciptest-selectionai-agentsregression-test-selectionpyrightcode-graphswe-benchai-coding-agentshydradbcode-as-graphs
-
Updated
Aug 23, 2026 - Python