Pith. sign in

Paper Citation Record · LEDGER

InterCode: Standardizing and Benchmarking Interactive Coding with Execution Feedback

As of 8 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 20 inbound Pith citation observations for arXiv:2306.14898.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2306.14898 v3

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 20 of 20 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 20 of 20 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T04:23:50.664728Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-06-29T11:23:20.895534Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 4e88b8a0-fefd-425c-8f95-4aee49b6af53 · inbound

OSWorld: Benchmarking Multimodal Agents for Open-Ended Tasks in Real Computer Environments cites this paper.

OSWorld: Benchmarking Multimodal Agents for Open-Ended Tasks in Real Computer Environments InterCode: Standardizing and Benchmarking Interactive Coding with Execution Feedback

Reference 58

Resolution
verified exact
arxiv_id, observed 2026-05-13T01:19:32.507302Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-13T01:19:32.406859Z digest=sha256:22b99a16d317988743bd02696b8ec2b4678c2456ac8dab2ee83ce8ab56668f2a

Observation c7fad0a8-726e-4417-be4f-a7eba076ae48 · inbound

Gemma 2: Improving Open Language Models at a Practical Size cites this paper.

Gemma 2: Improving Open Language Models at a Practical Size InterCode: Standardizing and Benchmarking Interactive Coding with Execution Feedback

Reference 56

Resolution
verified exact
arxiv_id, observed 2026-05-10T12:11:16.397576Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-05-10T12:11:16.326752Z digest=sha256:f435103e160c0d739b401d8ef4444f4b3341ce71a33fc6fd7498e9345e07b7bb

Observation 539ad4c5-a155-4bb9-aa70-e2f6f2e740f2 · inbound

Training Language Models to Self-Correct via Reinforcement Learning cites this paper.

Training Language Models to Self-Correct via Reinforcement Learning InterCode: Standardizing and Benchmarking Interactive Coding with Execution Feedback

Reference 147

Resolution
metadata mismatch
arxiv_id, observed 2026-05-17T12:04:10.693845Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-05-17T12:04:10.210508Z digest=sha256:2b70062de4be3bb3155dc43a1dacb42b7432e2ecce2a1f253c917e27183c6d8f

Observation a05762d5-d970-4558-a807-414ce6e31ab7 · inbound

OS-ATLAS: A Foundation Action Model for Generalist GUI Agents cites this paper.

OS-ATLAS: A Foundation Action Model for Generalist GUI Agents InterCode: Standardizing and Benchmarking Interactive Coding with Execution Feedback

Reference 53

Resolution
metadata mismatch
arxiv_id, observed 2026-05-13T09:29:27.725374Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-05-13T09:29:27.173784Z digest=sha256:99e31d807faa1d8917652d9b4f7c15fc475d011ba99ccc2fb04e34b777d72dc4

Observation fe4b2975-5b95-4c5c-9a92-7fd4b40436d2 · inbound

OPT-BENCH: Evaluating LLM Agent on Large-Scale Search Spaces Optimization Problems cites this paper.

OPT-BENCH: Evaluating LLM Agent on Large-Scale Search Spaces Optimization Problems InterCode: Standardizing and Benchmarking Interactive Coding with Execution Feedback

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-07T04:23:50.664728Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:23:50.664728Z digest=sha256:901e753a45c46fd73a1b094c16620531e39902bbe8b600e820aa1ec10cecdb8d

Observation dc6a2e0e-4271-4bd0-820f-bdcb8a7cdd84 · inbound

Gemini 2.5: Pushing the Frontier with Advanced Reasoning, Multimodality, Long Context, and Next Generation Agentic Capabilities cites this paper.

Gemini 2.5: Pushing the Frontier with Advanced Reasoning, Multimodality, Long Context, and Next Generation Agentic Capabilities InterCode: Standardizing and Benchmarking Interactive Coding with Execution Feedback

Reference 91

Resolution
verified exact
arxiv_id, observed 2026-05-19T05:52:07.763204Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-05-19T05:48:02.828938Z digest=sha256:029f983bd5377c3b63064efd011b75827c5aa7ac6f46d799de0422157286a808

Observation 421b26c3-468b-4eec-bfc1-9e04165ceaef · inbound

ExCyTIn-Bench: Evaluating LLM agents on Cyber Threat Investigation cites this paper.

ExCyTIn-Bench: Evaluating LLM agents on Cyber Threat Investigation InterCode: Standardizing and Benchmarking Interactive Coding with Execution Feedback

Reference 53

Resolution
verified exact
arxiv_id, observed 2026-05-19T04:42:04.731715Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-19T04:37:33.942379Z digest=sha256:6116004c47f03c56e564b09cb6d8667edae85514509cfc08ef4a0ba6de4edc8d

Observation 4f30a2d7-2449-48b5-a025-1fc9b2f9f650 · inbound

Running in CIRCLE? A Simple Benchmark for LLM Code Interpreter Security cites this paper.

Running in CIRCLE? A Simple Benchmark for LLM Code Interpreter Security InterCode: Standardizing and Benchmarking Interactive Coding with Execution Feedback

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-06T14:25:15.844022Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T14:25:15.844022Z digest=sha256:6ca2ee852371293c5953278e077711990a826c552aafa7f231925d1df1d926f3

Observation ff2ff9f9-012e-4dad-932e-cec5c038a972 · inbound

Feedback-Driven Execution for LLM-Based Binary Analysis cites this paper.

Feedback-Driven Execution for LLM-Based Binary Analysis InterCode: Standardizing and Benchmarking Interactive Coding with Execution Feedback

Reference 50

Resolution
metadata mismatch
arxiv_id, observed 2026-05-10T10:44:37.936889Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-10T10:40:32.133423Z digest=sha256:cff96c41da93ed1749983d3c21b57516a2c7b1235336844bf40cdf75290f7204

Observation 27b5ac5f-7a8c-49a3-a988-b6d093ff4ecb · inbound

Towards Optimal Agentic Architectures for Offensive Security Tasks cites this paper.

Towards Optimal Agentic Architectures for Offensive Security Tasks InterCode: Standardizing and Benchmarking Interactive Coding with Execution Feedback

Reference 17

Resolution
verified exact
arxiv_id, observed 2026-05-11T12:16:03.139155Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-10T04:02:04.359269Z digest=sha256:bb6687adcc9c330d3369c3f058ca61d4d0a5dac336ddecaa4b2c1cfe3d965b63

Observation 4842b0c6-41f6-44f9-9f67-e5f2811cbac0 · inbound

Meta-Tool: Efficient Few-Shot Tool Adaptation for Small Language Models cites this paper.

Meta-Tool: Efficient Few-Shot Tool Adaptation for Small Language Models InterCode: Standardizing and Benchmarking Interactive Coding with Execution Feedback

Reference 52

Resolution
verified exact
arxiv_id, observed 2026-05-10T00:54:48.809796Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-05-10T00:50:27.257888Z digest=sha256:31ae45e47e4ed945a289e9d12d6196277b9826632501e71cbdff75c98bf15853

Observation 4759661e-1d2d-4fcd-b639-63c929da966d · inbound

Toward Scalable Terminal Task Synthesis via Skill Graphs cites this paper.

Toward Scalable Terminal Task Synthesis via Skill Graphs InterCode: Standardizing and Benchmarking Interactive Coding with Execution Feedback

Reference 14

Resolution
verified exact
arxiv_id, observed 2026-05-11T23:51:16.938897Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-07T16:13:50.484665Z digest=sha256:db2403922d3ab4e54d529f464e4507bfe3284b6418e628000a29d7c034f8315e

Observation c5a1fe64-f8c3-442e-bf6b-15e7ce40cc5b · inbound

OPT-BENCH: Evaluating the Iterative Self-Optimization of LLM Agents in Large-Scale Search Spaces cites this paper.

OPT-BENCH: Evaluating the Iterative Self-Optimization of LLM Agents in Large-Scale Search Spaces InterCode: Standardizing and Benchmarking Interactive Coding with Execution Feedback

Reference 141

Resolution
metadata mismatch
arxiv_id, observed 2026-05-12T03:01:18.775762Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-05-12T02:57:15.521594Z digest=sha256:a03dd455f698919b28d7aa9eedd4595ccfbb71deecaa243e1f98f484506dcb3f

Observation bedc2437-7737-4586-b6fb-561fb2ff7552 · inbound

CrackMeBench: Binary Reverse Engineering for Agents cites this paper.

CrackMeBench: Binary Reverse Engineering for Agents InterCode: Standardizing and Benchmarking Interactive Coding with Execution Feedback

Reference 12

Resolution
metadata mismatch
arxiv_id, observed 2026-05-12T05:56:43.456149Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-05-12T04:44:13.078520Z digest=sha256:7682eabfca825f787eff6a308ec69a54e6e738e2547ee93ccb47672f2f263126

Observation cfefc68d-e3a0-4e4b-8aa8-a6f859724145 · inbound

A-ProS: Towards Reliable Autonomous Programming Through Multi-Model Feedback cites this paper.

A-ProS: Towards Reliable Autonomous Programming Through Multi-Model Feedback InterCode: Standardizing and Benchmarking Interactive Coding with Execution Feedback

Reference 34

Resolution
verified exact
arxiv_id, observed 2026-05-20T09:23:10.555911Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-20T09:22:06.285118Z digest=sha256:79fdcf72fca773201dc7633bf3c0848cd8e99e306015d6db155fdcb1bd48a0c0

Observation 06905fc6-0fae-4c00-9649-0a9350f6218b · inbound

unix-ctf: Procedural Environments for Unix-Competence Reinforcement Learning cites this paper.

unix-ctf: Procedural Environments for Unix-Competence Reinforcement Learning InterCode: Standardizing and Benchmarking Interactive Coding with Execution Feedback

Reference 14

Resolution
metadata mismatch
arxiv_id, observed 2026-06-29T11:23:20.897179Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-29T11:19:38.959705Z digest=sha256:d812ca8ba9d383e46e7148abe50c2f7e8e1f84c8c654ed1cef0a8e39c97dc386

Observation cde89d7d-a132-4ca7-81a8-df4fc26dbda3 · inbound

Token Reduction Is Not Cost Reduction cites this paper.

Token Reduction Is Not Cost Reduction InterCode: Standardizing and Benchmarking Interactive Coding with Execution Feedback

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-02T06:45:15.404937Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T06:45:15.404937Z digest=sha256:f1c86f1ec1f0af99ad70ca9441c198a945c2fea00327a1ea2462b081f84cd5b4

Observation b5f905ca-bc27-4185-9528-d5ddceb2e1d9 · inbound

Token Reduction Is Not Cost Reduction cites this paper.

Token Reduction Is Not Cost Reduction InterCode: Standardizing and Benchmarking Interactive Coding with Execution Feedback

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-04T04:23:32.473137Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T04:23:32.473137Z digest=sha256:07166af1824d96128cfe33127d9ac3c48216b436200275a7ef98fc34d40efeaf

Observation 523c8696-edfe-4ff8-9149-b511a4f7c8d4 · inbound

The Disruptive Impact of Large Language Models on Capture the Flag Competitions and the Path Toward Fair Play cites this paper.

The Disruptive Impact of Large Language Models on Capture the Flag Competitions and the Path Toward Fair Play InterCode: Standardizing and Benchmarking Interactive Coding with Execution Feedback

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-01T02:31:26.014099Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T02:31:26.014099Z digest=sha256:d1caa1409fb6d55ff230f91b402d55f158ed6cf1a22dc00fea0470f7316d667a

Observation e1d6829c-1f02-40f2-920f-7b5caa53f733 · inbound

Learning to Coordinate Symbolic Tools: LLM Agents for Verified Sum-of-Squares Certificates cites this paper.

Learning to Coordinate Symbolic Tools: LLM Agents for Verified Sum-of-Squares Certificates InterCode: Standardizing and Benchmarking Interactive Coding with Execution Feedback

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-04T00:48:33.449476Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T00:48:33.449476Z digest=sha256:210cf5db95d4be81a608c3832274fa4e198a8f035e20358c88f1338044834417