Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-12T11:39:33.837756Z
Paper Citation Record · LEDGER
As of 13 August 2026, this Paper Citation Record lists 44 of 44 outbound references and 2 inbound Pith citation observations for arXiv:2411.18019.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-12T11:39:33.837756Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-12T06:34:41.77262+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-02T16:27:43.489582Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-05-11T08:26:01.477245Z
44 of 44 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 09da401d-680c-41cb-9c88-4677228200fa · outbound
A Real-World Benchmark for Evaluating Fine-Grained Issue Solving Capabilities of Large Language Models Unresolved cited work
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 218b4b6e-83fb-454b-90ab-4f796936a9cd · outbound
A Real-World Benchmark for Evaluating Fine-Grained Issue Solving Capabilities of Large Language Models Program Synthesis with Large Language Models
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ff833103-d0a0-4aef-90a7-92706b999a49 · outbound
A Real-World Benchmark for Evaluating Fine-Grained Issue Solving Capabilities of Large Language Models Unresolved cited work
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 7c550b34-48bc-4c13-90f6-7edb3122a433 · outbound
A Real-World Benchmark for Evaluating Fine-Grained Issue Solving Capabilities of Large Language Models Evaluating Large Language Models Trained on Code
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c385cfc2-44c4-4d5b-aa5d-3efc4dc463ad · outbound
A Real-World Benchmark for Evaluating Fine-Grained Issue Solving Capabilities of Large Language Models Unresolved cited work
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f6e0fc25-e88b-492b-a1cc-8fcd40f42907 · outbound
A Real-World Benchmark for Evaluating Fine-Grained Issue Solving Capabilities of Large Language Models Unresolved cited work
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation d36b637b-25cd-40f7-a11c-f28f7f5f6516 · outbound
A Real-World Benchmark for Evaluating Fine-Grained Issue Solving Capabilities of Large Language Models Generalization-Enhanced Code Vulnerability Detection via Multi-Task Instruction Fine-Tuning
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0e32bc69-8746-42ce-84c1-3048914e325c · outbound
A Real-World Benchmark for Evaluating Fine-Grained Issue Solving Capabilities of Large Language Models Unresolved cited work
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 8993847b-e09b-4c02-a348-b535ba8a88ff · outbound
A Real-World Benchmark for Evaluating Fine-Grained Issue Solving Capabilities of Large Language Models Unresolved cited work
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3c90e457-6e5d-4901-8970-c3ae95b5f194 · outbound
A Real-World Benchmark for Evaluating Fine-Grained Issue Solving Capabilities of Large Language Models CodeApex: A Bilingual Programming Evaluation Benchmark for Large Language Models
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b66db029-2a51-42cc-9860-3601c0dda63f · outbound
A Real-World Benchmark for Evaluating Fine-Grained Issue Solving Capabilities of Large Language Models Unresolved cited work
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 603ebf6a-a91d-4a0e-8569-ee18c40581ee · outbound
A Real-World Benchmark for Evaluating Fine-Grained Issue Solving Capabilities of Large Language Models Unresolved cited work
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation f9094f8b-fd9f-4cbd-b51f-8f7bd58ae0bc · outbound
A Real-World Benchmark for Evaluating Fine-Grained Issue Solving Capabilities of Large Language Models Measuring Coding Challenge Competence With APPS
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3d4846ff-6186-436e-8bcb-25942094c996 · outbound
A Real-World Benchmark for Evaluating Fine-Grained Issue Solving Capabilities of Large Language Models Unresolved cited work
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation f91bd4fa-8d70-4197-8546-c6aa0eb04328 · outbound
A Real-World Benchmark for Evaluating Fine-Grained Issue Solving Capabilities of Large Language Models Unresolved cited work
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 138d75c1-91f8-4f0e-93b1-dfaa283e39ab · outbound
A Real-World Benchmark for Evaluating Fine-Grained Issue Solving Capabilities of Large Language Models Unresolved cited work
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 4e0c54d3-92cd-429a-bede-39ee02206601 · outbound
A Real-World Benchmark for Evaluating Fine-Grained Issue Solving Capabilities of Large Language Models SWE-bench: Can Language Models Resolve Real-World GitHub Issues?
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3cb517ee-ad6b-48d5-a850-253c1869d346 · outbound
A Real-World Benchmark for Evaluating Fine-Grained Issue Solving Capabilities of Large Language Models Unresolved cited work
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2841977a-9a7e-455b-b836-2c4e3972167a · outbound
A Real-World Benchmark for Evaluating Fine-Grained Issue Solving Capabilities of Large Language Models Unresolved cited work
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 7794c252-cdcd-4f02-9bc9-8412265b1c88 · outbound
A Real-World Benchmark for Evaluating Fine-Grained Issue Solving Capabilities of Large Language Models CS1QA: A Dataset for Assisting Code-based Question Answering in an Introductory Programming Course
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 69294333-50c2-4ef3-b636-1bf4817087f7 · outbound
A Real-World Benchmark for Evaluating Fine-Grained Issue Solving Capabilities of Large Language Models Unresolved cited work
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation b4c27d15-aa5c-44db-8fb6-0b7301f71394 · outbound
A Real-World Benchmark for Evaluating Fine-Grained Issue Solving Capabilities of Large Language Models Unresolved cited work
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 33acb935-17d1-4061-8383-46031e2b2343 · outbound
A Real-World Benchmark for Evaluating Fine-Grained Issue Solving Capabilities of Large Language Models CodeQA: A Question Answering Dataset for Source Code Comprehension
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b18bc12a-fdfd-4b29-9623-ea71967496e3 · outbound
A Real-World Benchmark for Evaluating Fine-Grained Issue Solving Capabilities of Large Language Models Large Language Model-Based Agents for Software Engineering: A Survey
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 31f7ff52-f071-4e90-a145-3178924804fd · outbound
A Real-World Benchmark for Evaluating Fine-Grained Issue Solving Capabilities of Large Language Models MarsCode Agent: AI-native Automated Bug Fixing
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b47db412-079e-494d-a814-8a89fa9c97d6 · outbound
A Real-World Benchmark for Evaluating Fine-Grained Issue Solving Capabilities of Large Language Models Alibaba LingmaAgent: Improving Automated Issue Resolution via Comprehensive Repository Exploration
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 66788ad5-5064-4a54-b1d9-5af72ea5cec7 · outbound
A Real-World Benchmark for Evaluating Fine-Grained Issue Solving Capabilities of Large Language Models Unresolved cited work
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 671f294d-ae94-453c-a163-5aeb427b7e48 · outbound
A Real-World Benchmark for Evaluating Fine-Grained Issue Solving Capabilities of Large Language Models Unresolved cited work
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation aa3ec18d-c639-4023-9d2c-99dc5375fb90 · outbound
A Real-World Benchmark for Evaluating Fine-Grained Issue Solving Capabilities of Large Language Models Unresolved cited work
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1fdc7f2c-5d5b-4eb1-9228-2795fcff54a8 · outbound
A Real-World Benchmark for Evaluating Fine-Grained Issue Solving Capabilities of Large Language Models In Proceedings of the 54th ACM Technical Symposium on Computer Science Education V
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 0133477d-ebf4-49d4-b6a6-4910dcec5908 · outbound
A Real-World Benchmark for Evaluating Fine-Grained Issue Solving Capabilities of Large Language Models StarCoder: may the source be with you!
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 27c3fe17-275b-4b4e-880c-c764f2c3589b · outbound
A Real-World Benchmark for Evaluating Fine-Grained Issue Solving Capabilities of Large Language Models Unresolved cited work
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f09e517f-a269-43da-a9fa-f52a74cc2dbd · outbound
A Real-World Benchmark for Evaluating Fine-Grained Issue Solving Capabilities of Large Language Models AgentFL: Scaling LLM-based Fault Localization to Project-Level Context
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 446a64cd-193c-4482-a7cc-e40d4b44be73 · outbound
A Real-World Benchmark for Evaluating Fine-Grained Issue Solving Capabilities of Large Language Models Unresolved cited work
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 2f37515f-e2c9-4d98-9dee-2af068903482 · outbound
A Real-World Benchmark for Evaluating Fine-Grained Issue Solving Capabilities of Large Language Models MAGIS: LLM-Based Multi-Agent Framework for GitHub Issue Resolution
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 45ce0934-704d-4f33-9a3a-323af91f0e8d · outbound
A Real-World Benchmark for Evaluating Fine-Grained Issue Solving Capabilities of Large Language Models Unresolved cited work
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 496d94a6-a0c0-4c82-a9a9-25c5ece1129e · outbound
A Real-World Benchmark for Evaluating Fine-Grained Issue Solving Capabilities of Large Language Models Unresolved cited work
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 51e64f52-0f11-41cc-b9b2-2539d36e359f · outbound
A Real-World Benchmark for Evaluating Fine-Grained Issue Solving Capabilities of Large Language Models Agentless: Demystifying LLM-based Software Engineering Agents
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a524a3c0-1beb-4cf6-ac80-750ab6937de7 · outbound
A Real-World Benchmark for Evaluating Fine-Grained Issue Solving Capabilities of Large Language Models Unresolved cited work
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation b1235a58-3ac0-4cf2-bb7c-6467c8b9fea3 · outbound
A Real-World Benchmark for Evaluating Fine-Grained Issue Solving Capabilities of Large Language Models Unresolved cited work
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 00923c45-29b0-42ac-85a4-417bf0225f83 · outbound
A Real-World Benchmark for Evaluating Fine-Grained Issue Solving Capabilities of Large Language Models AutoCodeRover: Autonomous Program Improvement
Reference 41
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1be7cee2-c66c-424e-95eb-752600efba8c · outbound
A Real-World Benchmark for Evaluating Fine-Grained Issue Solving Capabilities of Large Language Models SWE-agent: Agent-Computer Interfaces Enable Automated Software Engineering
Reference 42
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5c952517-67fa-4ecd-a0d6-1a4c14c51d23 · outbound
A Real-World Benchmark for Evaluating Fine-Grained Issue Solving Capabilities of Large Language Models Large Language Models for Software Engineering: A Systematic Literature Review
Reference 2023
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 24de3a1a-f3f7-46ff-8bf9-9a7ea1c3be3b · outbound
A Real-World Benchmark for Evaluating Fine-Grained Issue Solving Capabilities of Large Language Models In Proceedings of the 46th IEEE/ACM International Conference on Software Engineering
Reference 2024
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 399dd694-103d-440f-80a1-257581758b6a · inbound
Evaluating LLM Agents on Automated Software Analysis Tasks A Real-World Benchmark for Evaluating Fine-Grained Issue Solving Capabilities of Large Language Models
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation f9bd4156-3400-4ce0-b5f5-ffdad8f9f162 · inbound
Evaluating LLM Agents on Automated Software Analysis Tasks A Real-World Benchmark for Evaluating Fine-Grained Issue Solving Capabilities of Large Language Models
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.