Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 23 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 35 inbound Pith citation observations for arXiv:2504.08703.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-23T06:30:58.430688+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-15T21:32:07.773449Z
A source-named dated measurement, never combined with another source.
Source: pith, observed 2026-08-05T02:28:24.338817Z
0 of 0 outbound references displayed
External citation measurements
0
pith, observed 2026-08-05T02:28:24.338817Z
No outbound reference observations are available for this paper version.
Observation 0ec09632-f2d9-4c6a-95be-8947c0d99554 · inbound
MigrationBench: Repository-Level Code Migration Benchmark from Java 8 SWE-PolyBench: A multi-language benchmark for repository level evaluation of coding agents
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c6840061-8020-4f91-9307-9c76666be5f1 · inbound
CoRet: Improved Retriever for Code Editing SWE-PolyBench: A multi-language benchmark for repository level evaluation of coding agents
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8d5a436d-2c7d-4f6c-a0c4-cc0393dca03d · inbound
SemAgent: A Semantics Aware Program Repair Agent SWE-PolyBench: A multi-language benchmark for repository level evaluation of coding agents
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 89ce5fba-1f0f-4c4c-a4b2-0eb064c6aa2e · inbound
Is Your Automated Software Engineer Trustworthy? SWE-PolyBench: A multi-language benchmark for repository level evaluation of coding agents
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f959a0d6-0aac-4caa-8f88-4f6a70d79a16 · inbound
AI-Assisted Fixes to Code Review Comments at Scale SWE-PolyBench: A multi-language benchmark for repository level evaluation of coding agents
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9db39d24-7605-4fab-ab05-23178030fac3 · inbound
NoCode-bench: A Benchmark for Evaluating Natural Language-Driven Feature Addition SWE-PolyBench: A multi-language benchmark for repository level evaluation of coding agents
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5a38868c-0153-4821-b29e-ffea33ada3bb · inbound
RepoDebug: Repository-Level Multi-Task and Multi-Language Debugging Evaluation of Large Language Models SWE-PolyBench: A multi-language benchmark for repository level evaluation of coding agents
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 556750d0-1a52-4700-a53a-c74ffca308bc · inbound
Vulcan: Instance-specialized, Verifiable Systems Heuristics Through LLM-driven Search SWE-PolyBench: A multi-language benchmark for repository level evaluation of coding agents
Reference 83
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7275fef0-9400-493e-bb0f-73e710672053 · inbound
BeyondSWE: Can Current Code Agent Survive Beyond Single-Repo Bug Fixing? SWE-PolyBench: A multi-language benchmark for repository level evaluation of coding agents
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4de5c26f-12ea-44f2-a967-ffdc32a0d7ec · inbound
Reproduction Test Generation for Java SWE Issues SWE-PolyBench: A multi-language benchmark for repository level evaluation of coding agents
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 084b3ac7-c203-40f1-b9a1-b5cc75fc3470 · inbound
Reproduction Test Generation for Java SWE Issues SWE-PolyBench: A multi-language benchmark for repository level evaluation of coding agents
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation dd5f6290-f84e-4999-9947-6ed841d1a0eb · inbound
Constraint Decay: The Fragility of LLM Agents in Backend Code Generation SWE-PolyBench: A multi-language benchmark for repository level evaluation of coding agents
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 79e4c14f-be29-40e1-b775-334c6b6e7ec4 · inbound
SkillMaster: Toward Autonomous Skill Mastery in LLM Agents SWE-PolyBench: A multi-language benchmark for repository level evaluation of coding agents
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 0ce48223-aa1f-4691-8e5e-fed59c3bf9d9 · inbound
SkillMaster: Toward Autonomous Skill Mastery in LLM Agents SWE-PolyBench: A multi-language benchmark for repository level evaluation of coding agents
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation d2c908b6-d85d-4bed-ae64-dfdcbab49330 · inbound
Do Coding Agents Understand Least-Privilege Authorization? SWE-PolyBench: A multi-language benchmark for repository level evaluation of coding agents
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 7a29ded5-7b9f-49f1-89a8-621a29370407 · inbound
Beyond Summaries: Structure-Aware Labeling of Code Changes with Large Language Models SWE-PolyBench: A multi-language benchmark for repository level evaluation of coding agents
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation ef2aa17a-230b-4cf0-bbe1-411dee9dd333 · inbound
RepoMirage: Probing Repository Context Reasoning in Code Agents with Perturbations SWE-PolyBench: A multi-language benchmark for repository level evaluation of coding agents
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation be5bfb7f-fc18-41eb-91f2-77ad0d2cecc9 · inbound
HARP: Measuring Harm Amplification in Multi-Agent LLM Systems SWE-PolyBench: A multi-language benchmark for repository level evaluation of coding agents
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 4ce69601-6eb9-402b-b649-8167127c7905 · inbound
SWE-InfraBench: Evaluating Language Models on Cloud Infrastructure Code SWE-PolyBench: A multi-language benchmark for repository level evaluation of coding agents
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 40079cff-c810-4807-af2b-5e86041df481 · inbound
Projecting the Emerging Mindset of SWE Agent by Launching a Wild Code Understanding Journey SWE-PolyBench: A multi-language benchmark for repository level evaluation of coding agents
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 3db99ded-e4dc-4bbd-be11-530475791d76 · inbound
Code Isn't Memory: A Structural Codebase Index Inside a Coding Agent SWE-PolyBench: A multi-language benchmark for repository level evaluation of coding agents
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 2f60a935-f5ec-46a0-a1ec-b17d58d41293 · inbound
Lingering Authority: Revocable Resource-and-Effect Capabilities for Coding Agents SWE-PolyBench: A multi-language benchmark for repository level evaluation of coding agents
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation a0e49560-a412-4918-ac6c-83696027dce1 · inbound
Unlocking Model Potentials Through Adaptive Multi-Agent Scaffolding for Efficient Issue Resolution SWE-PolyBench: A multi-language benchmark for repository level evaluation of coding agents
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation a4370db8-2b61-4139-983d-3fd729037edb · inbound
LLVM-Bench: Benchmarking and Advancing Large Language Models for LLVM Compiler Issue Resolution SWE-PolyBench: A multi-language benchmark for repository level evaluation of coding agents
Reference 50
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation d6b33905-2d98-4e07-9a77-ff3abf94e719 · inbound
SWE-Doctor: Guiding Software Engineering Agents with Runtime Diagnosis from Multi-Faceted Bug Reproduction Tests SWE-PolyBench: A multi-language benchmark for repository level evaluation of coding agents
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 1a25496d-0ef0-4a49-8363-9c8639d30965 · inbound
What Resolve Rate Hides: Trajectory Structure Diagnostics for Coding Agents SWE-PolyBench: A multi-language benchmark for repository level evaluation of coding agents
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 38dcf282-46d8-40dc-b7b4-5cbd9d787ffa · inbound
From Execution to Education: A Bloom-Aligned Framework for Measuring Educational Control in LLMs SWE-PolyBench: A multi-language benchmark for repository level evaluation of coding agents
Reference 149
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 868d74f9-c2a1-4362-a2f8-dc85f2d96a19 · inbound
When Does Restricting a Coding Agent to execute_code Help? A Regime $\times$ Agent-Design Ablation SWE-PolyBench: A multi-language benchmark for repository level evaluation of coding agents
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c6dbbf03-e191-4648-b221-062e31cd8bef · inbound
Bespoke Visual Assistance: What and How do Blind and Low-Vision People Create with Agentic Programming? SWE-PolyBench: A multi-language benchmark for repository level evaluation of coding agents
Reference 88
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 93ad891a-740d-4833-80ac-59b502d4e4f4 · inbound
ExplainBench: Evaluating Code Explanations from Agents SWE-PolyBench: A multi-language benchmark for repository level evaluation of coding agents
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6f6eecbf-26b4-4811-83cd-9892d8dafca8 · inbound
SWE-NFI: Studying and Benchmarking Coding Agents for Non-Functional Improvements SWE-PolyBench: A multi-language benchmark for repository level evaluation of coding agents
Reference 57
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation dbfa0fc3-3702-4796-b194-091c90ce97f7 · inbound
PAIChecker: Uncovering and Checking PR-Issue Misalignment in SWE-Bench-Like Benchmarks SWE-PolyBench: A multi-language benchmark for repository level evaluation of coding agents
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c876ea63-a64d-4a70-b5fb-2e4fe5bf4988 · inbound
PAIChecker: Uncovering and Checking PR-Issue Misalignment in SWE-Bench-Like Benchmarks SWE-PolyBench: A multi-language benchmark for repository level evaluation of coding agents
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a70a37d2-dabb-4c9b-9b66-d9a51647561f · inbound
SWE-Bench ProMax: Benchmarking Agents on Large-Scale Multilingual Code Refactoring SWE-PolyBench: A multi-language benchmark for repository level evaluation of coding agents
Reference 48
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation babcfd37-37e3-40c3-8eda-c32d9e406937 · inbound
One Recipe, Many Harnesses: What Self-Evolution Encodes Across Languages and Models SWE-PolyBench: A multi-language benchmark for repository level evaluation of coding agents
Reference 2025
Source-reported events for the cited work
Unavailable: canonical work link unavailable.