Archer automates agentic code review for LLVM optimizations and reports finding semantic bugs in 21% of recent open PRs and 11% of closed PRs.
Codellm-Devkit: A framework for contextualizing code LLMs with program analysis insights
4 Pith papers cite this work, alongside 8 external citations. Polarity classification is still indexing.
representative citing papers
Vigil deploys a proactive agent for full on-call lifecycle support with autonomous self-improvement from human-resolved cases.
PRAXIS combines LLM-driven structured traversal of service dependency graphs and hammock-block program dependence graphs to improve root-cause analysis accuracy by up to 6.3x while cutting token consumption by 5.3x on 30 real-world cloud incidents.
TestPrune minimizes regression test suites to improve bug reproduction and patch validation in LLM-based agentic repair pipelines, delivering 6-13% relative gains on SWE-Bench benchmarks at low API cost.
citing papers explorer
-
Archer: Towards Agentic Review for Compiler Optimizations
Archer automates agentic code review for LLVM optimizations and reports finding semantic bugs in 21% of recent open PRs and 11% of closed PRs.
-
Help Without Being Asked: A Deployed Proactive Agent System for On-Call Support with Continuous Self-Improvement
Vigil deploys a proactive agent for full on-call lifecycle support with autonomous self-improvement from human-resolved cases.
-
PRAXIS: Integrating Program Analysis with Observability for Root-Cause Analysis
PRAXIS combines LLM-driven structured traversal of service dependency graphs and hammock-block program dependence graphs to improve root-cause analysis accuracy by up to 6.3x while cutting token consumption by 5.3x on 30 real-world cloud incidents.
-
Can Old Tests Do New Tricks for Resolving SWE Issues?
TestPrune minimizes regression test suites to improve bug reproduction and patch validation in LLM-based agentic repair pipelines, delivering 6-13% relative gains on SWE-Bench benchmarks at low API cost.