Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 14 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 41 inbound Pith citation observations for arXiv:2410.06992.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-14T06:32:32.682623+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-14T04:39:14.638892Z
A source-named dated measurement, never combined with another source.
Source: pith, observed 2026-07-09T06:06:01.699658Z
0 of 0 outbound references displayed
External citation measurements
No source-named external measurement is stored.
No outbound reference observations are available for this paper version.
Observation 734c1bc4-4623-40cb-aa05-49b7d4d376e7 · inbound
Are Large Language Models Memorizing Bug Benchmarks? SWE-Bench+: Enhanced Coding Benchmark for LLMs
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b0720d43-ead2-4c3b-8ad8-dc7370652551 · inbound
TDD-Bench Verified: Can LLMs Generate Tests for Issues Before They Get Resolved? SWE-Bench+: Enhanced Coding Benchmark for LLMs
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 33aeee14-9ed4-4008-ba10-896630427d4f · inbound
Empirical Research on Utilizing LLM-based Agents for Automated Bug Fixing via LangGraph SWE-Bench+: Enhanced Coding Benchmark for LLMs
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d6fd010e-2f4e-401d-ae27-e2b6d8badf8d · inbound
Survey on Evaluation of LLM-based Agents SWE-Bench+: Enhanced Coding Benchmark for LLMs
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 03c1fc2a-f737-4294-88ae-31b39d749740 · inbound
AI Scientists Fail Without Strong Implementation Capability SWE-Bench+: Enhanced Coding Benchmark for LLMs
Reference 2025
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 28f50129-f104-4548-8edb-974f1bb26935 · inbound
UTBoost: Rigorous Evaluation of Coding Agents on SWE-Bench SWE-Bench+: Enhanced Coding Benchmark for LLMs
Reference 2024
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 813acb66-9098-400d-ad9c-6e5b2a189a9b · inbound
Evolutionary Perspectives on the Evaluation of LLM-Based AI Agents: A Comprehensive Survey SWE-Bench+: Enhanced Coding Benchmark for LLMs
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d597f088-f028-4633-961e-c630e8acdb65 · inbound
SWE-MERA: A Dynamic Benchmark for Agenticly Evaluating Large Language Models on Software Engineering Tasks SWE-Bench+: Enhanced Coding Benchmark for LLMs
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5fe3e2bd-be30-4584-8c04-edcb7807487d · inbound
A Survey of Self-Evolving Agents: What, When, How, and Where to Evolve on the Path to Artificial Super Intelligence SWE-Bench+: Enhanced Coding Benchmark for LLMs
Reference 108
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 1c4d3d22-d861-4f52-a7f0-60a2169c301b · inbound
What You See Is What It Does: A Structural Pattern for Legible Software SWE-Bench+: Enhanced Coding Benchmark for LLMs
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c44fd470-a130-4890-8f22-92769b81ba68 · inbound
Agentic Software Engineering: Foundational Pillars and a Research Roadmap SWE-Bench+: Enhanced Coding Benchmark for LLMs
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b58572c5-3108-493a-9244-811664d4d9a7 · inbound
SWE-Bench Pro: Can AI Agents Solve Long-Horizon Software Engineering Tasks? SWE-Bench+: Enhanced Coding Benchmark for LLMs
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 3d751a8e-f0a8-45e3-9085-a24626676f65 · inbound
Estimating the Empowerment of Language Model Agents SWE-Bench+: Enhanced Coding Benchmark for LLMs
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4e515c7b-4519-4917-a171-49dd09df8e32 · inbound
A Retrieval-Augmented Generation Approach to Extracting Algorithmic Logic from Neural Networks SWE-Bench+: Enhanced Coding Benchmark for LLMs
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation dd23f142-01f3-45c8-b6ae-3612b4efa103 · inbound
AgentCE-Bench: Agent Configurable Evaluation with Scalable Horizons and Controllable Difficulty under Lightweight Environments SWE-Bench+: Enhanced Coding Benchmark for LLMs
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 78123d6e-1e37-40cf-883e-f4213a472ff7 · inbound
REAgent: Requirement-Driven LLM Agents for Software Issue Resolution SWE-Bench+: Enhanced Coding Benchmark for LLMs
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 9d788d73-1faf-4d00-a682-5796b9fbe2a3 · inbound
CityRAG: Stepping Into a City via Spatially-Grounded Video Generation SWE-Bench+: Enhanced Coding Benchmark for LLMs
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 8644411e-088a-42ff-8e80-a56acb2fa379 · inbound
PlayCoder: Making LLM-Generated GUI Code Playable SWE-Bench+: Enhanced Coding Benchmark for LLMs
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 514c8e86-c50e-4e64-8c78-2c2856b3ce60 · inbound
The Conversations Beneath the Code: Triadic Data for Long-Horizon Software Engineering Agents SWE-Bench+: Enhanced Coding Benchmark for LLMs
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation b3ec8b0d-554a-4710-89d4-434401d9793e · inbound
Coding Agents Don't Know When to Act SWE-Bench+: Enhanced Coding Benchmark for LLMs
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 32d1af8b-1f24-4007-a172-c56beb2cce83 · inbound
From Patches to Trajectories: Privileged Process Supervision for Software-Engineering Agents SWE-Bench+: Enhanced Coding Benchmark for LLMs
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 91c5ad04-f5b0-404d-811d-91d711ec4459 · inbound
Anchor: Mitigating Artifact Drift in Agent Benchmark Generation SWE-Bench+: Enhanced Coding Benchmark for LLMs
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation c5f9e042-b8c3-45b1-aea9-4fe307b18575 · inbound
Code-QA-Bench: Separating Code Reasoning from Documentation Memorization in Repository-Level QA SWE-Bench+: Enhanced Coding Benchmark for LLMs
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 5013e746-5aff-4c80-a489-d021058eac91 · inbound
I-WebGenBench : Evaluating Interactivity in LLM-Generated Scientific Web Applications SWE-Bench+: Enhanced Coding Benchmark for LLMs
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation bcfeee33-2178-4232-aa02-7999abfd4782 · inbound
TensorBench: Benchmarking Coding Agents on a Compiler-Based Tensor Framework SWE-Bench+: Enhanced Coding Benchmark for LLMs
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 74a6e7ec-23b4-46e1-86ea-3b71d5864453 · inbound
Human Oversight and Overload: Two Hidden and Costly Burdens of AI-Assisted Software Engineering SWE-Bench+: Enhanced Coding Benchmark for LLMs
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 96715c5b-8349-4eba-9f9d-686873e19d27 · inbound
Position: Coding Benchmarks Are Misaligned with Agentic Software Engineering SWE-Bench+: Enhanced Coding Benchmark for LLMs
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation c8044ef7-40fa-4f3b-84f6-8e453a2a5e20 · inbound
Position: Coding Benchmarks Are Misaligned with Agentic Software Engineering SWE-Bench+: Enhanced Coding Benchmark for LLMs
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e17a204b-9cf9-4428-a277-273c667f27df · inbound
Breaking the Solver Bottleneck: Training Task Generators at the Learnable Frontier SWE-Bench+: Enhanced Coding Benchmark for LLMs
Reference 50
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 97eb9eae-2037-4c27-a592-2107b09e0fb7 · inbound
Code Isn't Memory: A Structural Codebase Index Inside a Coding Agent SWE-Bench+: Enhanced Coding Benchmark for LLMs
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation d34adb41-8833-431e-8aca-e387f96dc721 · inbound
Helpful or Harmful? Evaluating LLM-Assisted Vulnerability Patching via a Human Study SWE-Bench+: Enhanced Coding Benchmark for LLMs
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 21d2878d-09f4-4eec-995d-416e3da6ae80 · inbound
SWE-Router: Routing in Multi-turn Agentic Software Engineering Tasks SWE-Bench+: Enhanced Coding Benchmark for LLMs
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation a818805b-ade7-4898-bed3-a26264c355fc · inbound
RuBench: A Repository-Level Agentic Coding Benchmark with Natively Authored Russian Task Specifications SWE-Bench+: Enhanced Coding Benchmark for LLMs
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 7f097b81-7b4e-4618-9fa7-f37081d6ea79 · inbound
RuBench: A Repository-Level Agentic Coding Benchmark with Natively Authored Russian Task Specifications SWE-Bench+: Enhanced Coding Benchmark for LLMs
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6ac850aa-93ec-4f35-af31-facf3bda70a1 · inbound
What Makes a Good Bug Report for an AI Agent? SWE-Bench+: Enhanced Coding Benchmark for LLMs
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation b0533e7d-7473-4b5e-bacb-f0655eb0f325 · inbound
Deep Interaction: An Efficient Human-AI Interaction Method for Large Reasoning Models SWE-Bench+: Enhanced Coding Benchmark for LLMs
Reference 2024
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9ad5f78a-feae-458e-8159-6525e23e059e · inbound
PAIChecker: Uncovering and Checking PR-Issue Misalignment in SWE-Bench-Like Benchmarks SWE-Bench+: Enhanced Coding Benchmark for LLMs
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4d0f7d26-67ee-4a82-911d-41e3cf95e93a · inbound
PAIChecker: Uncovering and Checking PR-Issue Misalignment in SWE-Bench-Like Benchmarks SWE-Bench+: Enhanced Coding Benchmark for LLMs
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 64ba4e39-b21a-4b0a-9c41-5cddd61a9a18 · inbound
Active-SWE: Benchmarking Coding Agents for Proactive Bug Fixing without Issue Reports SWE-Bench+: Enhanced Coding Benchmark for LLMs
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b4660dad-0725-4a6e-ae3e-a9009c2266ea · inbound
The Horizon Gap: Planning, Memory, Execution, Training, and Evaluation for Long-Horizon LLM Agents SWE-Bench+: Enhanced Coding Benchmark for LLMs
Reference 103
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 220ce96b-cc93-4b25-b1a9-999fa60dbb0f · inbound
Tangent: An Empirical Study of Testing Practices for LLM-Based Agent Applications SWE-Bench+: Enhanced Coding Benchmark for LLMs
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.