Pith. sign in

Paper Citation Record · LEDGER

Reasoning Runtime Behavior of a Program with LLM: How Far Are We?

As of 16 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 15 inbound Pith citation observations for arXiv:2403.16437.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2403.16437 v3

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 15 of 15 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-16T06:30:59.297886+00:00

measured 15 of 15 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-16T12:39:09.573317Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-20T13:18:18.504080Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 0fbb1997-8f7a-4b8c-a04f-e8b41c096c37 · inbound

Revisit Self-Debugging with Self-Generated Tests for Code Generation cites this paper.

Revisit Self-Debugging with Self-Generated Tests for Code Generation Reasoning Runtime Behavior of a Program with LLM: How Far Are We?

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-10T16:51:39.184683Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T16:51:39.184683Z digest=sha256:3a8a913903d0dddc238055cbe06faaa3bec7a29b21f5a9c16a334a0c7232404e

Observation 99d810e9-d44a-4571-8dc1-02e5fd678b7d · inbound

Revisit Self-Debugging with Self-Generated Tests for Code Generation cites this paper.

Revisit Self-Debugging with Self-Generated Tests for Code Generation Reasoning Runtime Behavior of a Program with LLM: How Far Are We?

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-10T16:51:39.189567Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T16:51:39.189567Z digest=sha256:23c43eae4aab733623fe6f3b4649991c353e195749a2dbd32181fda75db4f2f9

Observation ab520caa-90c8-46c1-ad45-70795c519053 · inbound

Correctness Assessment of Code Generated by Large Language Models Using Internal Representations cites this paper.

Correctness Assessment of Code Generated by Large Language Models Using Internal Representations Reasoning Runtime Behavior of a Program with LLM: How Far Are We?

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-10T16:41:30.968570Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T16:41:30.968570Z digest=sha256:de6d1fde8e24a048190105179265afed15821fd45a5dc1dac40d9bb3f3a4c20b

Observation ca356ac8-4503-48c0-8ee3-fde4e711400c · inbound

Themisto: Jupyter-Based Runtime Benchmark cites this paper.

Themisto: Jupyter-Based Runtime Benchmark Reasoning Runtime Behavior of a Program with LLM: How Far Are We?

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-08-16T12:39:09.573317Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T12:39:09.573317Z digest=sha256:2b66b70883ab14b4341aeb6f45a90d1a0ceddcffb7db33c36070d1cbefe01651

Observation 836c8ef8-0d64-4947-bf9b-9f06c91e59d6 · inbound

Large Language Models for Validating Network Protocol Parsers cites this paper.

Large Language Models for Validating Network Protocol Parsers Reasoning Runtime Behavior of a Program with LLM: How Far Are We?

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-16T12:11:29.371954Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T12:11:29.371954Z digest=sha256:7c85c0897705f3c9d83f5f025340636bd408567f8583c822c42847fc67dd6768

Observation 429d2b99-726b-41f9-8a8e-d868dd008394 · inbound

CodeReasoner: Enhancing the Code Reasoning Ability with Reinforcement Learning cites this paper.

CodeReasoner: Enhancing the Code Reasoning Ability with Reinforcement Learning Reasoning Runtime Behavior of a Program with LLM: How Far Are We?

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-06T14:50:27.845007Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T14:50:27.845007Z digest=sha256:5eb8fef8d0f20066ed3cd2b24b1c7c1273671fb31e25e952b71682afd46b2cf8

Observation 112c9392-b6ca-49d3-a9d7-c49ade4da725 · inbound

Is "Knowing It's Malicious Enough?" Evaluating LLMs for Fine-Grained Malware Behavior Auditing cites this paper.

Is "Knowing It's Malicious Enough?" Evaluating LLMs for Fine-Grained Malware Behavior Auditing Reasoning Runtime Behavior of a Program with LLM: How Far Are We?

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-15T15:56:17.432853Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T15:56:17.432853Z digest=sha256:307a060329b2ad2104e8785217bd77af581e1c6c274c5ebe8bfab591710a3b2c

Observation 8abf2aad-fd56-48d2-a9ae-ffd92cae8f1d · inbound

Evaluating Code Reasoning Abilities of Large Language Models Under Real-World Settings cites this paper.

Evaluating Code Reasoning Abilities of Large Language Models Under Real-World Settings Reasoning Runtime Behavior of a Program with LLM: How Far Are We?

Reference 9

Resolution
verified exact
arxiv_id, observed 2026-05-16T21:28:34.242076Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-05-16T21:23:44.762007Z digest=sha256:78a3259e817912b0d4371030476f86874285d5ce7af964db1458f85dd01a9795

Observation 67104567-29bb-4455-9500-76d5bf907c95 · inbound

PrismaDV: Automated Task-Aware Data Unit Test Generation cites this paper.

PrismaDV: Automated Task-Aware Data Unit Test Generation Reasoning Runtime Behavior of a Program with LLM: How Far Are We?

Reference 6

Resolution
verified exact
arxiv_id, observed 2026-05-09T22:49:16.092795Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-05-09T22:14:30.829159Z digest=sha256:c95513a0a81bd9b62281d24803c69fbc5c2a8c11b80e8200605ea5a89667d3e3

Observation c13ff555-23ad-45f9-b8d6-f5f9071277f2 · inbound

StepCodeReasoner: Aligning Code Reasoning with Stepwise Execution Traces via Reinforcement Learning cites this paper.

StepCodeReasoner: Aligning Code Reasoning with Stepwise Execution Traces via Reinforcement Learning Reasoning Runtime Behavior of a Program with LLM: How Far Are We?

Reference 22

Resolution
metadata mismatch
arxiv_id, observed 2026-05-13T05:32:20.277543Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-05-13T05:27:37.521421Z digest=sha256:6eeeb5c5b1c6e623a595ed3ce476f37df4742b9778aa5fef6d1c22d88075417a

Observation f1da4b3a-9d38-4ae7-be12-7ef3628a7a09 · inbound

Veritas: Grounding LLM Agents for Reliable Vulnerability Reasoning over Stripped Binaries cites this paper.

Veritas: Grounding LLM Agents for Reliable Vulnerability Reasoning over Stripped Binaries Reasoning Runtime Behavior of a Program with LLM: How Far Are We?

Reference 13

Resolution
verified exact
arxiv_id, observed 2026-05-15T14:05:54.699476Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-05-15T14:04:16.443346Z digest=sha256:38c4da9093dbc22aa8e2730bd08fff4ffaa3c91b58939109e6200df256618cf2

Observation 49b215d5-e978-4967-997c-f696f5bc0e0d · inbound

Veritas: Grounding LLM Agents for Reliable Vulnerability Reasoning over Stripped Binaries cites this paper.

Veritas: Grounding LLM Agents for Reliable Vulnerability Reasoning over Stripped Binaries Reasoning Runtime Behavior of a Program with LLM: How Far Are We?

Reference 13

Resolution
unresolved
no resolver link, observed 2026-07-12T16:46:36.842198Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T16:46:36.842198Z digest=sha256:3f9723b9bc1ab3291ed9f1370a9d6d27f96c3602d68d0efbfd12414329142fdc

Observation 634c1555-1267-4590-a966-9671b5a6c608 · inbound

Veritas: Grounding LLM Agents for Reliable Vulnerability Reasoning over Stripped Binaries cites this paper.

Veritas: Grounding LLM Agents for Reliable Vulnerability Reasoning over Stripped Binaries Reasoning Runtime Behavior of a Program with LLM: How Far Are We?

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-02T14:03:35.157446Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T14:03:35.157446Z digest=sha256:acdaaa499337345999138e1877e69dbd613e846a291203c168efb7fb191d03ff

Observation efa2aa85-e696-43f6-ba4d-2ec73ccb5503 · inbound

Enhancing the Code Reasoning Capabilities of LLMs via Consistency-based Reinforcement Learning cites this paper.

Enhancing the Code Reasoning Capabilities of LLMs via Consistency-based Reinforcement Learning Reasoning Runtime Behavior of a Program with LLM: How Far Are We?

Reference 4

Resolution
verified exact
arxiv_id, observed 2026-05-20T13:18:18.506006Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-05-20T13:13:51.081597Z digest=sha256:785c8a99da2a88102c1e36b8a71dbd4e490d578e48539e5b2b7db024086073f9

Observation e5c2710d-5e12-459e-bd55-419dfed68c47 · inbound

RepoReasoner: Evaluating Repository-Level Code Reasoning Ability of Long-Context Language Models cites this paper.

RepoReasoner: Evaluating Repository-Level Code Reasoning Ability of Long-Context Language Models Reasoning Runtime Behavior of a Program with LLM: How Far Are We?

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-01T00:57:33.349547Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T00:57:33.349547Z digest=sha256:10cec49217b0ebb057db2fbfdf060b790d7056a1eb13f96c2502dc78c253a0f1