Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-16T01:04:43.459105Z
Paper Citation Record · LEDGER
As of 20 August 2026, this Paper Citation Record lists 42 of 42 outbound references and 7 inbound Pith citation observations for arXiv:2505.02133.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-16T01:04:43.459105Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-19T06:32:44.657259+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-07T14:15:48.502829Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-07-01T20:26:12.298078Z
42 of 42 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 63bb2c79-f556-49ea-a6f7-8a66d6ddcd94 · outbound
Enhancing LLM Code Generation: A Systematic Evaluation of Multi-Agent Collaboration and Runtime Debugging for Improved Accuracy, Reliability, and Latency GPT-4o System Card,
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 891f2155-9151-4dc2-ab21-4d87868b876f · outbound
Enhancing LLM Code Generation: A Systematic Evaluation of Multi-Agent Collaboration and Runtime Debugging for Improved Accuracy, Reliability, and Latency The Claude 3 Model Family: Opus, Sonnet, Haiku Anthropic
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 566dc2b3-0fb3-4fa0-9b48-e731c902a45c · outbound
Enhancing LLM Code Generation: A Systematic Evaluation of Multi-Agent Collaboration and Runtime Debugging for Improved Accuracy, Reliability, and Latency The Llama 3 Herd of Models
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 34f46a74-908c-4b6c-8ef4-e2089dbbd57c · outbound
Enhancing LLM Code Generation: A Systematic Evaluation of Multi-Agent Collaboration and Runtime Debugging for Improved Accuracy, Reliability, and Latency Gemini: A Family of Highly Capable Multimodal Models
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d6cc477e-051d-41b4-a8a8-a5972107493f · outbound
Enhancing LLM Code Generation: A Systematic Evaluation of Multi-Agent Collaboration and Runtime Debugging for Improved Accuracy, Reliability, and Latency Gemma 2: Improving Open Language Models at a Practical Size,
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 3cf27a4f-3245-4c05-be8a-9e65418cc5e5 · outbound
Enhancing LLM Code Generation: A Systematic Evaluation of Multi-Agent Collaboration and Runtime Debugging for Improved Accuracy, Reliability, and Latency LLaMA: Open and Efficient Foundation Language Models
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 356a5fc6-1917-4645-aa4d-33c03899eb1f · outbound
Enhancing LLM Code Generation: A Systematic Evaluation of Multi-Agent Collaboration and Runtime Debugging for Improved Accuracy, Reliability, and Latency Phi-3 Technical Report: A Highly Capable Language Model Locally on Your Phone
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2f83eb6b-2293-489b-8372-f75c999c899d · outbound
Enhancing LLM Code Generation: A Systematic Evaluation of Multi-Agent Collaboration and Runtime Debugging for Improved Accuracy, Reliability, and Latency Qwen2 Technical Report,
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation ed9ec64f-bb83-450b-a6ce-dc0377c002c2 · outbound
Enhancing LLM Code Generation: A Systematic Evaluation of Multi-Agent Collaboration and Runtime Debugging for Improved Accuracy, Reliability, and Latency Mixtral of Experts,
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 6a816ab2-4977-4a41-9046-4cf92011ae6b · outbound
Enhancing LLM Code Generation: A Systematic Evaluation of Multi-Agent Collaboration and Runtime Debugging for Improved Accuracy, Reliability, and Latency Mistral 7B
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8bd75353-218e-477c-8631-14995790576a · outbound
Enhancing LLM Code Generation: A Systematic Evaluation of Multi-Agent Collaboration and Runtime Debugging for Improved Accuracy, Reliability, and Latency DeepSeek-V3 Technical Report
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f0fc007e-beb8-4ab2-8814-b75a364472a5 · outbound
Enhancing LLM Code Generation: A Systematic Evaluation of Multi-Agent Collaboration and Runtime Debugging for Improved Accuracy, Reliability, and Latency Zephyr: Direct Distillation of LM Alignment
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b8314428-cd69-4830-ab02-a5e74bfdba4f · outbound
Enhancing LLM Code Generation: A Systematic Evaluation of Multi-Agent Collaboration and Runtime Debugging for Improved Accuracy, Reliability, and Latency Evaluating Large Language Models Trained on Code
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8efe2e37-41fc-4496-80eb-b096aca127c8 · outbound
Enhancing LLM Code Generation: A Systematic Evaluation of Multi-Agent Collaboration and Runtime Debugging for Improved Accuracy, Reliability, and Latency From LLMs to LLM-based Agents for Software Engineering: A Survey of Current, Challenges and Future
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 13c4f10b-ffe2-4cc2-9b95-ee16f70fd48a · outbound
Enhancing LLM Code Generation: A Systematic Evaluation of Multi-Agent Collaboration and Runtime Debugging for Improved Accuracy, Reliability, and Latency Unifying the Perspectives of NLP and Software Engineering: A Survey on Language Models for Code
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ed483d6e-8357-407d-b9d9-70472e7a24f4 · outbound
Enhancing LLM Code Generation: A Systematic Evaluation of Multi-Agent Collaboration and Runtime Debugging for Improved Accuracy, Reliability, and Latency AgentCoder: Multi-Agent-based Code Generation with Iterative Testing and Optimisation
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c3152481-f5b3-4957-8eb9-9c49009869ff · outbound
Enhancing LLM Code Generation: A Systematic Evaluation of Multi-Agent Collaboration and Runtime Debugging for Improved Accuracy, Reliability, and Latency MapCoder: Multi-Agent Code Generation for Competitive Problem Solving
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 397b354b-bcfc-42e0-a65e-f567584cbd13 · outbound
Enhancing LLM Code Generation: A Systematic Evaluation of Multi-Agent Collaboration and Runtime Debugging for Improved Accuracy, Reliability, and Latency Scaling Large Language Model-based Multi-Agent Collaboration
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f205114c-7806-4330-9011-b1b6b66be203 · outbound
Enhancing LLM Code Generation: A Systematic Evaluation of Multi-Agent Collaboration and Runtime Debugging for Improved Accuracy, Reliability, and Latency ChatDev: Communicative Agents for Software Development
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 76be9be7-25f4-47af-a821-01f241a3546b · outbound
Enhancing LLM Code Generation: A Systematic Evaluation of Multi-Agent Collaboration and Runtime Debugging for Improved Accuracy, Reliability, and Latency MetaGPT: Meta Programming for A Multi-Agent Collaborative Framework,
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation a21010ec-3fb1-4770-8208-e1fcef4692ce · outbound
Enhancing LLM Code Generation: A Systematic Evaluation of Multi-Agent Collaboration and Runtime Debugging for Improved Accuracy, Reliability, and Latency Self-collaboration Code Generation via ChatGPT
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6c38a6e5-4dc2-4946-8027-ad502f73829b · outbound
Enhancing LLM Code Generation: A Systematic Evaluation of Multi-Agent Collaboration and Runtime Debugging for Improved Accuracy, Reliability, and Latency CY- CLE: Learning to Self-Refine the Code Genera- tion,
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e5d92926-ae7b-4961-aa24-17b5e79640b3 · outbound
Enhancing LLM Code Generation: A Systematic Evaluation of Multi-Agent Collaboration and Runtime Debugging for Improved Accuracy, Reliability, and Latency Debug like a Hu- man: A Large Language Model Debugger via Ver- ifying Runtime Execution Step-by-step,
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 5bbc7345-4b66-4432-9907-3920a25ce4e7 · outbound
Enhancing LLM Code Generation: A Systematic Evaluation of Multi-Agent Collaboration and Runtime Debugging for Improved Accuracy, Reliability, and Latency Leveraging Print Debugging to Improve Code Gen- eration in Large Language Models,
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation e5ea0a8e-e10e-46fa-b1b4-1e235c1b6bec · outbound
Enhancing LLM Code Generation: A Systematic Evaluation of Multi-Agent Collaboration and Runtime Debugging for Improved Accuracy, Reliability, and Latency RGD: Multi-LLM Based Agent Debugger via Refinement and Generation Guidance
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1d3b70fd-2043-4a3b-b87b-43f975e65c9c · outbound
Enhancing LLM Code Generation: A Systematic Evaluation of Multi-Agent Collaboration and Runtime Debugging for Improved Accuracy, Reliability, and Latency From Code to Correctness: Closing the Last Mile of Code Gen- eration with Hierarchical Debugging,
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 48863728-1fc6-4104-b23d-878133734bb2 · outbound
Enhancing LLM Code Generation: A Systematic Evaluation of Multi-Agent Collaboration and Runtime Debugging for Improved Accuracy, Reliability, and Latency SOEN-101: Code Generation by Emulating Software Process Models Using Large Language Model Agents
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 04b168da-97fc-41e5-b35e-d44deba62306 · outbound
Enhancing LLM Code Generation: A Systematic Evaluation of Multi-Agent Collaboration and Runtime Debugging for Improved Accuracy, Reliability, and Latency ReflectionCoder: Learning from Reflection Sequence for Enhanced One-off Code Generation
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ce530737-14de-4e1a-9d60-7924b50b05db · outbound
Enhancing LLM Code Generation: A Systematic Evaluation of Multi-Agent Collaboration and Runtime Debugging for Improved Accuracy, Reliability, and Latency Teaching Large Language Models to Self-Debug,
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation a0042412-a25e-4f03-971f-99b8954eee62 · outbound
Enhancing LLM Code Generation: A Systematic Evaluation of Multi-Agent Collaboration and Runtime Debugging for Improved Accuracy, Reliability, and Latency Program Synthesis with Large Language Models
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation dc747d94-56fa-43eb-bcd3-e24efdf0b5cc · outbound
Enhancing LLM Code Generation: A Systematic Evaluation of Multi-Agent Collaboration and Runtime Debugging for Improved Accuracy, Reliability, and Latency Is Your Code Generated by ChatGPT Really Cor- rect? Rigorous Evaluation of Large Language Mod- els for Code Generation
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation b74e3a94-0f41-4441-a020-d4bfe7c69838 · outbound
Enhancing LLM Code Generation: A Systematic Evaluation of Multi-Agent Collaboration and Runtime Debugging for Improved Accuracy, Reliability, and Latency BLOOM: A 176B-Parameter Open-Access Multilingual Language Model
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4c15df1f-98a2-470d-b476-8660ae97f09a · outbound
Enhancing LLM Code Generation: A Systematic Evaluation of Multi-Agent Collaboration and Runtime Debugging for Improved Accuracy, Reliability, and Latency PaLM: Scaling Language Modeling with Pathways,
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 87f81ae2-f6df-43d1-bf1b-ee63a9359b47 · outbound
Enhancing LLM Code Generation: A Systematic Evaluation of Multi-Agent Collaboration and Runtime Debugging for Improved Accuracy, Reliability, and Latency Qwen Technical Report,
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation a5cb156e-f217-4a1c-a89a-9c90034eb762 · outbound
Enhancing LLM Code Generation: A Systematic Evaluation of Multi-Agent Collaboration and Runtime Debugging for Improved Accuracy, Reliability, and Latency DeepSeekMoE: Towards Ultimate Expert Specialization in Mixture-of-Experts Language Models
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 36b8f8ac-2c08-4ec1-8412-7eccf762652b · outbound
Enhancing LLM Code Generation: A Systematic Evaluation of Multi-Agent Collaboration and Runtime Debugging for Improved Accuracy, Reliability, and Latency Code Llama: Open Foundation Models for Code
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 37e5dbd0-3c1b-4c40-8042-9fb50580b369 · outbound
Enhancing LLM Code Generation: A Systematic Evaluation of Multi-Agent Collaboration and Runtime Debugging for Improved Accuracy, Reliability, and Latency WizardCoder: Empowering Code Large Language Models with Evol-Instruct
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 87ff476a-64d7-49c2-9ad7-c0aecf2585c8 · outbound
Enhancing LLM Code Generation: A Systematic Evaluation of Multi-Agent Collaboration and Runtime Debugging for Improved Accuracy, Reliability, and Latency StarCoder: may the source be with you!
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 39909963-21db-4165-8d7f-221b5d1efbf5 · outbound
Enhancing LLM Code Generation: A Systematic Evaluation of Multi-Agent Collaboration and Runtime Debugging for Improved Accuracy, Reliability, and Latency Magicoder: Empowering Code Generation with OSS-Instruct
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3399c2de-76b1-4782-a35a-ff29a92b24bb · outbound
Enhancing LLM Code Generation: A Systematic Evaluation of Multi-Agent Collaboration and Runtime Debugging for Improved Accuracy, Reliability, and Latency OpenCodeInterpreter: Integrating Code Generation with Execution and Refinement
Reference 41
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 15b5655c-2c1c-4b8a-9e5e-d27cce6b770b · outbound
Enhancing LLM Code Generation: A Systematic Evaluation of Multi-Agent Collaboration and Runtime Debugging for Improved Accuracy, Reliability, and Latency StarCoder 2 and The Stack v2: The Next Generation
Reference 42
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation dee68f70-8c12-4d26-a275-425b2f933eac · outbound
Enhancing LLM Code Generation: A Systematic Evaluation of Multi-Agent Collaboration and Runtime Debugging for Improved Accuracy, Reliability, and Latency Reflexion: Language Agents with Verbal Reinforcement Learning
Reference 43
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0793178c-9d13-43b2-9189-2a21611f57ec · inbound
Vibe Coding vs. Agentic Coding: Fundamentals and Practical Implications of Agentic AI Enhancing LLM Code Generation: A Systematic Evaluation of Multi-Agent Collaboration and Runtime Debugging for Improved Accuracy, Reliability, and Latency
Reference 176
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0f8e2ab5-3c98-4d47-9e65-b73712e18143 · inbound
Position Paper: Programming Language Techniques for Bridging LLM Code Generation Semantic Gaps Enhancing LLM Code Generation: A Systematic Evaluation of Multi-Agent Collaboration and Runtime Debugging for Improved Accuracy, Reliability, and Latency
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 349621a9-5105-4119-ad89-854ca063cbbb · inbound
WildCode Revisited: A Comprehensive Empirical Study on the Security of LLM-Generated Code Enhancing LLM Code Generation: A Systematic Evaluation of Multi-Agent Collaboration and Runtime Debugging for Improved Accuracy, Reliability, and Latency
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8ac82032-2948-4021-94d7-9ffae6761838 · inbound
Cascaded Code Editing: Large-Small Model Collaboration for Effective and Efficient Code Editing Enhancing LLM Code Generation: A Systematic Evaluation of Multi-Agent Collaboration and Runtime Debugging for Improved Accuracy, Reliability, and Latency
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 6d8f5d4b-db41-4aa0-9144-2998e43f7c0a · inbound
The Infinite Mutation Engine? Measuring Polymorphism in LLM-Generated Offensive Code Enhancing LLM Code Generation: A Systematic Evaluation of Multi-Agent Collaboration and Runtime Debugging for Improved Accuracy, Reliability, and Latency
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 914550fb-d923-462a-b161-56bb1db492a3 · inbound
The Infinite Mutation Engine? Measuring Polymorphism in LLM-Generated Offensive Code Enhancing LLM Code Generation: A Systematic Evaluation of Multi-Agent Collaboration and Runtime Debugging for Improved Accuracy, Reliability, and Latency
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation de17b9d3-59f7-4d1a-aee2-20ecf94db210 · inbound
How Generation Architecture Shapes Code Complexity in Multi-Agent LLM Systems: A Paired Study on HumanEval Enhancing LLM Code Generation: A Systematic Evaluation of Multi-Agent Collaboration and Runtime Debugging for Improved Accuracy, Reliability, and Latency
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.