Agent triage, measured: does this edit need a human, or can the AI handle it? We measure whether the graph under the agent can reach the tests that guard a change (0.42 pooled, 7 repos) and ship the fail-closed gate that routes unproven maps to human verification. CLI · MCP · SARIF · in-engine on HydraDB.
hackathon reproducible-research mcp static-analysis code-analysis developer-tools graph-database call-graph scip test-selection ai-agents regression-test-selection pyright code-graph swe-bench ai-coding-agents hydradb code-as-graphs
-
Updated
Aug 23, 2026 - Python