Pith. sign in

Paper Citation Record · LEDGER

Debug like a Human: A Large Language Model Debugger via Verifying Runtime Execution Step-by-step

As of 8 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 23 inbound Pith citation observations for arXiv:2402.16906.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2402.16906 v6

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 23 of 23 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 23 of 23 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T14:49:35.726058Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

3
arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 04a7b415-3d41-4651-9b90-e8370852942a · inbound

Specification-Driven Code Translation Powered by Large Language Models: How Far Are We? cites this paper.

Specification-Driven Code Translation Powered by Large Language Models: How Far Are We? Debug like a Human: A Large Language Model Debugger via Verifying Runtime Execution Step-by-step

Reference 16

Resolution
verified exact
arxiv_id, observed 2026-05-23T07:35:28.768708Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-23T07:35:19.736921Z digest=sha256:3d481ff2269b9df9285a5915939efae613847c5a38f7f4631b987a949b6c0b77

Observation 6439f9ff-47aa-407f-b472-50c46a5dbb26 · inbound

MARCO: Meta-Reflection with Cross-Referencing for Code Reasoning cites this paper.

MARCO: Meta-Reflection with Cross-Referencing for Code Reasoning Debug like a Human: A Large Language Model Debugger via Verifying Runtime Execution Step-by-step

Reference 80

Resolution
unresolved
no resolver link, observed 2026-08-07T14:49:35.726058Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:49:35.726058Z digest=sha256:ca9d429127668a3a633c53c4f7350d526f133f5e9a52cbffbb9f86999fb8e467

Observation e48c0abb-38b3-46e7-a054-709324835ec9 · inbound

Enhancing LLM-Based Code Generation with Complexity Metrics: A Feedback-Driven Approach cites this paper.

Enhancing LLM-Based Code Generation with Complexity Metrics: A Feedback-Driven Approach Debug like a Human: A Large Language Model Debugger via Verifying Runtime Execution Step-by-step

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-07T12:42:48.385369Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:42:48.385369Z digest=sha256:7fcfb85537a45bd3c0e699e6104fb29e7e68a8595cf92f91e883fd2ecea5bbd0

Observation f1a9f071-1258-4a48-8b0f-b0dc9ec1bbdf · inbound

Xolver: Multi-Agent Reasoning with Holistic Experience Learning Just Like an Olympiad Team cites this paper.

Xolver: Multi-Agent Reasoning with Holistic Experience Learning Just Like an Olympiad Team Debug like a Human: A Large Language Model Debugger via Verifying Runtime Execution Step-by-step

Reference 70

Resolution
unresolved
no resolver link, observed 2026-08-07T00:24:14.998105Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:24:14.998105Z digest=sha256:d5cc947ff823b5739a6c007bd22c123292e80856c9d6031926b1bed9dffd1490

Observation 234bcb10-d3a6-47f2-bc4d-fe64c78c9cb2 · inbound

Measuring and Augmenting Large Language Models for Solving Capture-the-Flag Challenges cites this paper.

Measuring and Augmenting Large Language Models for Solving Capture-the-Flag Challenges Debug like a Human: A Large Language Model Debugger via Verifying Runtime Execution Step-by-step

Reference 98

Resolution
unresolved
no resolver link, observed 2026-08-06T23:35:06.211001Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:35:06.211001Z digest=sha256:8dd5e3e7e26f7f06863cb5c996dc4e2b80143e7b9f1455a199a32fcf7a4583b2

Observation 880d12be-9dc4-422d-9236-c69685b58a62 · inbound

CodeReasoner: Enhancing the Code Reasoning Ability with Reinforcement Learning cites this paper.

CodeReasoner: Enhancing the Code Reasoning Ability with Reinforcement Learning Debug like a Human: A Large Language Model Debugger via Verifying Runtime Execution Step-by-step

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-06T14:50:28.075565Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T14:50:28.075565Z digest=sha256:22846fbd8a1297e69a40b54c6ed4d1ba3d4ac25b98402746b8d7256b275335c9

Observation cb38483e-6e4a-4f33-898d-fb433cae4d56 · inbound

Running in CIRCLE? A Simple Benchmark for LLM Code Interpreter Security cites this paper.

Running in CIRCLE? A Simple Benchmark for LLM Code Interpreter Security Debug like a Human: A Large Language Model Debugger via Verifying Runtime Execution Step-by-step

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-06T14:25:15.861206Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T14:25:15.861206Z digest=sha256:f62759138b0d292c1a387400c0cd44f3ddc8d2e32acbc8bef4655049652c4b5a

Observation 5e17f278-5263-41ab-abdf-1c10041c4853 · inbound

HLSDebugger: Identification and Correction of Logic Bugs in HLS Code with LLM Solutions cites this paper.

HLSDebugger: Identification and Correction of Logic Bugs in HLS Code with LLM Solutions Debug like a Human: A Large Language Model Debugger via Verifying Runtime Execution Step-by-step

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-06T12:46:40.769641Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T12:46:40.769641Z digest=sha256:e7fce9ec5cf0ea527391f7f953cd1b052d0920613d68812755fa8591d7cafc84

Observation 138b9b54-5a36-400d-9835-0d97c8b54655 · inbound

Let's Revise Step-by-Step: A Unified Local Search Framework for Code Generation with LLMs cites this paper.

Let's Revise Step-by-Step: A Unified Local Search Framework for Code Generation with LLMs Debug like a Human: A Large Language Model Debugger via Verifying Runtime Execution Step-by-step

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-05T22:09:46.518940Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T22:09:46.518940Z digest=sha256:ff4d11f36a250e072b8081bf9c56fc4f459fe6494c154c3e5e4839671dd4a94a

Observation e1d02a06-62f1-441f-a554-2ccef80e06e9 · inbound

ReCode: Improving LLM-based Code Repair with Fine-Grained Retrieval-Augmented Generation cites this paper.

ReCode: Improving LLM-based Code Repair with Fine-Grained Retrieval-Augmented Generation Debug like a Human: A Large Language Model Debugger via Verifying Runtime Execution Step-by-step

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-05T11:39:14.083993Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T11:39:14.083993Z digest=sha256:e873fb354cd7b2e26250daca95841ec292433469ef0a2f3bc27d91be1885fa22

Observation e7b0c769-9aa3-437d-9e2f-29051d296fb2 · inbound

In Line with Context: Repository-Level Code Generation via Context Inlining cites this paper.

In Line with Context: Repository-Level Code Generation via Context Inlining Debug like a Human: A Large Language Model Debugger via Verifying Runtime Execution Step-by-step

Reference 73

Resolution
verified exact
arxiv_id, observed 2026-05-16T18:08:12.808019Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-16T18:04:56.915339Z digest=sha256:8278848a3beec98c762a90f18f75f002d1b52dd109adcfe0d9abbc682caa519f

Observation e6e73bf9-4857-48a0-bdd4-8c8853180c78 · inbound

Debug2Fix: Can Interactive Debugging Help Coding Agents Fix More Bugs? cites this paper.

Debug2Fix: Can Interactive Debugging Help Coding Agents Fix More Bugs? Debug like a Human: A Large Language Model Debugger via Verifying Runtime Execution Step-by-step

Reference 19

Resolution
verified exact
arxiv_id, observed 2026-05-15T20:20:17.842508Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-15T20:17:52.610119Z digest=sha256:f2cd050615ae8b5f6cde66bf3fb6a98ffb63d03afe285535021cffa82d6b10a1

Observation 423b9b4e-3673-43b2-8a38-ecff1cf43351 · inbound

Detection Time Distribution Predicted Using Absorbing Boundary Conditions and Imaginary Potentials cites this paper.

Detection Time Distribution Predicted Using Absorbing Boundary Conditions and Imaginary Potentials Debug like a Human: A Large Language Model Debugger via Verifying Runtime Execution Step-by-step

Reference 28

Resolution
unresolved
no resolver link, observed 2026-07-13T20:28:38.179448Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-13T20:28:38.179448Z digest=sha256:393972bfaf2157e1bbc88896f2f63851e8719774e850b4b09298a2f1fb8261b9

Observation 647f40af-d011-4e14-a6bc-f4ae8690cf05 · inbound

Dynamic analysis enhances issue resolution cites this paper.

Dynamic analysis enhances issue resolution Debug like a Human: A Large Language Model Debugger via Verifying Runtime Execution Step-by-step

Reference 28

Resolution
verified exact
arxiv_id, observed 2026-05-15T00:48:24.895620Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-15T00:45:55.460545Z digest=sha256:59a6b2267d0d672293fafaf754dc150a6e5e53485b403eec1b0953dd3f3915af

Observation 4e80d75e-b4bc-487a-a323-cd5751583e95 · inbound

AdverMCTS: Combating Pseudo-Correctness in Code Generation via Adversarial Monte Carlo Tree Search cites this paper.

AdverMCTS: Combating Pseudo-Correctness in Code Generation via Adversarial Monte Carlo Tree Search Debug like a Human: A Large Language Model Debugger via Verifying Runtime Execution Step-by-step

Reference 60

Resolution
verified exact
arxiv_id, observed 2026-05-11T08:40:57.597712Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-05-10T16:35:16.056397Z digest=sha256:796fb0f6218f171b44fbb63d0dcfd4d997d789da7a20e2d9567ba5e5d0f0ad03

Observation 3123ca8f-b908-45bc-b825-786b8f5e1c02 · inbound

How Many Tries Does It Take? Iterative Self-Repair in LLM Code Generation Across Model Scales and Benchmarks cites this paper.

How Many Tries Does It Take? Iterative Self-Repair in LLM Code Generation Across Model Scales and Benchmarks Debug like a Human: A Large Language Model Debugger via Verifying Runtime Execution Step-by-step

Reference 24

Resolution
verified exact
arxiv_id, observed 2026-05-11T09:06:00.395039Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-10T16:14:52.614444Z digest=sha256:e3e0ba3ea92fad6892fc47f3b32d8f4c2b11238ff6b143ad22cb3ab1c32e00bd

Observation 151509f8-ab1c-4d19-952d-34cb0ad2a0c3 · inbound

A Comprehensive Survey on Agent Skills: Taxonomy, Techniques, and Applications cites this paper.

A Comprehensive Survey on Agent Skills: Taxonomy, Techniques, and Applications Debug like a Human: A Large Language Model Debugger via Verifying Runtime Execution Step-by-step

Reference 52

Resolution
verified exact
arxiv_id, observed 2026-05-11T04:20:57.021636Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-11T01:47:39.926540Z digest=sha256:f96609f376c47d855f67e00fd9e399f22ae0625de83c41f86d48adb2d3a59089

Observation 9bbb7300-3ad0-4605-b27a-e1643094d53b · inbound

A Comprehensive Survey on Agent Skills: Taxonomy, Techniques, and Applications cites this paper.

A Comprehensive Survey on Agent Skills: Taxonomy, Techniques, and Applications Debug like a Human: A Large Language Model Debugger via Verifying Runtime Execution Step-by-step

Reference 52

Resolution
verified exact
arxiv_id, observed 2026-05-20T23:19:15.120133Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-20T23:15:44.550045Z digest=sha256:50d76e5112dd2436cae7b67bd1d9605797b22e1e05332faa027a58b90cee2b32

Observation 08a2cae3-f3c1-48f9-a53e-f567bd35fa8e · inbound

A Comprehensive Survey on Agent Skills: Taxonomy, Techniques, and Applications cites this paper.

A Comprehensive Survey on Agent Skills: Taxonomy, Techniques, and Applications Debug like a Human: A Large Language Model Debugger via Verifying Runtime Execution Step-by-step

Reference 50

Resolution
verified exact
arxiv_id, observed 2026-06-30T23:25:07.438777Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-30T23:23:42.883286Z digest=sha256:06d960b106952405652fd5dfede982303f7d068d801c174981549b42f49e410e

Observation d59e8a9a-00b7-4d77-b013-f328f93e4991 · inbound

StepCodeReasoner: Aligning Code Reasoning with Stepwise Execution Traces via Reinforcement Learning cites this paper.

StepCodeReasoner: Aligning Code Reasoning with Stepwise Execution Traces via Reinforcement Learning Debug like a Human: A Large Language Model Debugger via Verifying Runtime Execution Step-by-step

Reference 39

Resolution
metadata mismatch
arxiv_id, observed 2026-05-13T05:32:20.209348Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-05-13T05:27:37.521421Z digest=sha256:67db3c0940ce3f5a207268e2a5b64eda3ef3dc45322971cd29cca643acadb1c9

Observation ee9575bd-67af-4c33-97d7-552400cb5328 · inbound

Prompt Optimization for LLM Code Generation via Reinforcement Learning cites this paper.

Prompt Optimization for LLM Code Generation via Reinforcement Learning Debug like a Human: A Large Language Model Debugger via Verifying Runtime Execution Step-by-step

Reference 38

Resolution
metadata mismatch
arxiv_id, observed 2026-05-20T08:53:10.295304Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-20T08:49:36.986452Z digest=sha256:54cd6f7c88cc30385a49ddc4f95571bf4be5e5d3e64d97934860e8ec2293e113

Observation 9b9dc526-1674-4f4b-8bc2-571b80a7c23a · inbound

How Generation Architecture Shapes Code Complexity in Multi-Agent LLM Systems: A Paired Study on HumanEval cites this paper.

How Generation Architecture Shapes Code Complexity in Multi-Agent LLM Systems: A Paired Study on HumanEval Debug like a Human: A Large Language Model Debugger via Verifying Runtime Execution Step-by-step

Reference 13

Resolution
verified exact
arxiv_id, observed 2026-07-01T20:26:12.311180Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-28T21:14:01.070601Z digest=sha256:e83dec33cd3672d3696136e1124f10fa1ac6f2613de7f9950842f2d98b7a8647

Observation 09042a0b-9425-4295-bc19-d8f1fb10914f · inbound

SrDetection: A Self-Referential Framework for Data Leakage Detection in Code Large Language Models cites this paper.

SrDetection: A Self-Referential Framework for Data Leakage Detection in Code Large Language Models Debug like a Human: A Large Language Model Debugger via Verifying Runtime Execution Step-by-step

Reference 26

Resolution
verified exact
arxiv_id, observed 2026-06-30T06:24:18.354146Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-06-30T06:20:02.046020Z digest=sha256:0fef65068beb600e3294127d2fcf52e01ca251607dd4c61527ce666a8f84f693