Pith. sign in

Paper Citation Record · LEDGER

SandboxEval: Towards Securing Test Environment for Untrusted Code

As of 17 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 8 inbound Pith citation observations for arXiv:2504.00018.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2504.00018 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 8 of 8 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-17T06:30:58.91139+00:00

measured 8 of 8 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-06T14:25:15.831596Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-07-11T02:47:50.407349Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 757a066a-c992-4e16-941d-852acb3c799c · inbound

LLM Agents Are the Antidote to Walled Gardens cites this paper.

LLM Agents Are the Antidote to Walled Gardens SandboxEval: Towards Securing Test Environment for Untrusted Code

Reference 72

Resolution
verified exact
arxiv_id, observed 2026-05-22T00:44:29.338949Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-05-22T00:41:19.750928Z digest=sha256:468264200306e79fcae8e57db2a8e038b4479d8bbc39c0a2b90c9b2f5d620338

Observation 4b7dc307-f238-4f90-a136-5d74533e6078 · inbound

Running in CIRCLE? A Simple Benchmark for LLM Code Interpreter Security cites this paper.

Running in CIRCLE? A Simple Benchmark for LLM Code Interpreter Security SandboxEval: Towards Securing Test Environment for Untrusted Code

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-06T14:25:15.831596Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T14:25:15.831596Z digest=sha256:b6f6e8d9a048ef60f68e671f68286b0516951a919bf5fd60ce12116f21da27c0

Observation 4c2ee53f-ba9e-469c-a30a-566d3ea05e3f · inbound

Agentic AI Security: Threats, Defenses, Evaluation, and Open Challenges cites this paper.

Agentic AI Security: Threats, Defenses, Evaluation, and Open Challenges SandboxEval: Towards Securing Test Environment for Untrusted Code

Reference 211

Resolution
metadata mismatch
arxiv_id, observed 2026-05-18T03:42:21.992547Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-05-18T03:42:10.703369Z digest=sha256:39ee3cdb9d7ed089aff96ed0a656ff53ceaec5b94b8552d7312c74636cdf19b8

Observation f6e49bfa-0f43-4eb9-8754-e76c1c33aafd · inbound

OOM-RL: Out-of-Money Reinforcement Learning Market-Driven Alignment for LLM-Based Multi-Agent Systems cites this paper.

OOM-RL: Out-of-Money Reinforcement Learning Market-Driven Alignment for LLM-Based Multi-Agent Systems SandboxEval: Towards Securing Test Environment for Untrusted Code

Reference 16

Resolution
verified exact
arxiv_id, observed 2026-05-11T08:25:58.773890Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-05-10T16:40:14.537670Z digest=sha256:fd852006472004d177e10856cab5a793ab4e9c27e8affe1d63a3d7792ed5a979

Observation ebda6e72-cbbd-465d-91fd-0b42df582b41 · inbound

PYTHALAB-MERA: Validation-Grounded Memory, Retrieval, and Acceptance Control for Frozen-LLM Coding Agents cites this paper.

PYTHALAB-MERA: Validation-Grounded Memory, Retrieval, and Acceptance Control for Frozen-LLM Coding Agents SandboxEval: Towards Securing Test Environment for Untrusted Code

Reference 59

Resolution
verified exact
arxiv_id, observed 2026-05-12T08:21:26.006427Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-05-12T01:12:37.970638Z digest=sha256:b320e1cfa68f26ce0fdeb06aeea74ff602bb52e9a250d35a37b4001aa01d005d

Observation 40d3a104-0017-49a6-8a6e-9865b2295591 · inbound

Do Coding Agents Understand Least-Privilege Authorization? cites this paper.

Do Coding Agents Understand Least-Privilege Authorization? SandboxEval: Towards Securing Test Environment for Untrusted Code

Reference 37

Resolution
metadata mismatch
arxiv_id, observed 2026-05-19T16:37:40.136547Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-05-19T16:34:14.379419Z digest=sha256:66a7304e32fb4cbd76263bc035de0612cfa65199f2cb4bed2a40eb5a95d892ac

Observation a03ea466-55b2-47b0-8f75-01d42dd90425 · inbound

The Balkanization of Execution-Security Research for AI Coding Agents: Isolation, Access Control, and Time-of-Check-to-Time-of-Use Vulnerabilities cites this paper.

The Balkanization of Execution-Security Research for AI Coding Agents: Isolation, Access Control, and Time-of-Check-to-Time-of-Use Vulnerabilities SandboxEval: Towards Securing Test Environment for Untrusted Code

Reference 22

Resolution
metadata mismatch
local_arxiv, observed 2026-07-11T02:47:50.429009Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-07-11T02:41:24.813416Z digest=sha256:d7937f0ce75e6536e7a26e1dfb9ed7f36f1d19785741e4fcff6ac605e9cf9667

Observation 9132506f-f561-45ec-9d9e-6ae928114e8c · inbound

Agent Security Needs Redefinition through a Holistic Framework cites this paper.

Agent Security Needs Redefinition through a Holistic Framework SandboxEval: Towards Securing Test Environment for Untrusted Code

Reference 228

Resolution
unresolved
no resolver link, observed 2026-08-01T06:04:46.410322Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T06:04:46.410322Z digest=sha256:723325a1ec52463f3c69a4a0d48229a4d43d75971e39917e69e52d49efba3b6a