Pith. sign in

Paper Citation Record · LEDGER

AgentGym: Evolving Large Language Model-based Agents across Diverse Environments

As of 4 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 17 inbound Pith citation observations for arXiv:2406.04151.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2406.04151 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 17 of 17 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-04T06:34:03.388597+00:00

measured 17 of 17 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-04T15:25:49.227255Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-07-08T04:04:29.283127Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation a4192876-0ea8-429f-9da7-a155f72cff9c · inbound

Agent Q: Advanced Reasoning and Learning for Autonomous AI Agents cites this paper.

Agent Q: Advanced Reasoning and Learning for Autonomous AI Agents AgentGym: Evolving Large Language Model-based Agents across Diverse Environments

Reference 71

Resolution
verified exact
arxiv_id, observed 2026-05-20T09:42:04.322581Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-20T09:41:59.979595Z digest=sha256:526212720ed4ea59a654f4d2f1ea32e01d0943fe57fd8ab52198701f7bdb7512

Observation 8d393d12-5068-41cb-a44d-1b946843cb2f · inbound

Towards Large Reasoning Models: A Survey of Reinforced Reasoning with Large Language Models cites this paper.

Towards Large Reasoning Models: A Survey of Reinforced Reasoning with Large Language Models AgentGym: Evolving Large Language Model-based Agents across Diverse Environments

Reference 166

Resolution
verified exact
arxiv_id, observed 2026-05-15T21:20:59.369509Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-15T21:20:59.128986Z digest=sha256:8680de3378d652ffeb55dd9da4128530333ae4fb23bb723ff508678ae232e9ad

Observation 633d89ad-8345-4656-9680-37d45a7c6df4 · inbound

GLM-4.5: Agentic, Reasoning, and Coding (ARC) Foundation Models cites this paper.

GLM-4.5: Agentic, Reasoning, and Coding (ARC) Foundation Models AgentGym: Evolving Large Language Model-based Agents across Diverse Environments

Reference 46

Resolution
verified exact
arxiv_id, observed 2026-05-11T17:50:08.538990Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-11T17:50:08.399160Z digest=sha256:4954558603ff18aa7454f21faeaa70cc140f2aa186d8c3b09144d4b3bc2f2792

Observation 9e9ec0d0-78c0-440f-8b6a-f8f10e505721 · inbound

From Refusal to Recovery: A Control-Theoretic Approach to Generative AI Guardrails cites this paper.

From Refusal to Recovery: A Control-Theoretic Approach to Generative AI Guardrails AgentGym: Evolving Large Language Model-based Agents across Diverse Environments

Reference 63

Resolution
verified exact
arxiv_id, observed 2026-05-21T20:44:21.982282Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-21T20:42:40.823721Z digest=sha256:255859fb6ca28388289955e46c641609445f26103b8a7f2eca675e22cefd64c4

Observation b60a2987-9a79-4afe-b54d-839ea9a8336e · inbound

SEAL: Synergistic Co-Evolution of Agents and Learning Environments cites this paper.

SEAL: Synergistic Co-Evolution of Agents and Learning Environments AgentGym: Evolving Large Language Model-based Agents across Diverse Environments

Reference 45

Resolution
verified exact
arxiv_id, observed 2026-06-30T13:44:40.906170Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-06-30T13:38:10.466713Z digest=sha256:4b6fad8371fb7cb9521ab49e611be1bcbe91f899dd8fbc5209e0de5027b8dd67

Observation 503b0cf1-cd55-4649-ba5a-033f3ca27251 · inbound

Test-Time Deep Thinking to Explore Implicit Rules cites this paper.

Test-Time Deep Thinking to Explore Implicit Rules AgentGym: Evolving Large Language Model-based Agents across Diverse Environments

Reference 37

Resolution
verified exact
arxiv_id, observed 2026-06-30T11:54:38.202000Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-06-30T11:52:15.163893Z digest=sha256:aa05c86d3d1d4ed02656e73861544421cb46fbc7103d692a1057ac74d1921af5

Observation 84f171e6-fd0e-40b2-bce1-9a0d91612a4f · inbound

Learning to Adapt: Self-Improving Web Agent via Cognitive-Aware Exploration cites this paper.

Learning to Adapt: Self-Improving Web Agent via Cognitive-Aware Exploration AgentGym: Evolving Large Language Model-based Agents across Diverse Environments

Reference 28

Resolution
verified exact
arxiv_id, observed 2026-07-01T19:36:08.811784Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-06-28T22:18:15.244033Z digest=sha256:7105ac93c78ee3143d48bf0b1ef386293c6ad25eff717483b95c7f2787cca743

Observation cde52fbb-e5c7-4822-9c73-a604b322cb8c · inbound

COMAP: Co-Evolving World Models and Agent Policies for LLM Agents cites this paper.

COMAP: Co-Evolving World Models and Agent Policies for LLM Agents AgentGym: Evolving Large Language Model-based Agents across Diverse Environments

Reference 58

Resolution
metadata mismatch
arxiv_id, observed 2026-07-01T23:16:24.914814Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-06-28T14:27:50.260308Z digest=sha256:e0f51fd62e431c8cf317f0eeaeabb2e7f40a56ff74977dea6a0a37de3a3f1098

Observation adea69dc-cc9a-40ab-b884-67c397419fd3 · inbound

A Scalable Approach to Evaluating Moral Sensitivity in LLMs cites this paper.

A Scalable Approach to Evaluating Moral Sensitivity in LLMs AgentGym: Evolving Large Language Model-based Agents across Diverse Environments

Reference 176

Resolution
unresolved
no resolver link, observed 2026-07-12T05:44:33.099337Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-12T05:44:33.099337Z digest=sha256:0679b23b1fb93857febb915ad378bafabf7c5d45aa3efd5d1d3d79262620267e

Observation a778e4cb-40b6-4358-a5f8-9d10192347b5 · inbound

Doomed from the Start: Early Abort of LLM Agent Episodes via a Recall-Controlled Probe Cascade cites this paper.

Doomed from the Start: Early Abort of LLM Agent Episodes via a Recall-Controlled Probe Cascade AgentGym: Evolving Large Language Model-based Agents across Diverse Environments

Reference 24

Resolution
metadata mismatch
local_arxiv, observed 2026-07-08T04:04:29.284362Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-07-08T03:54:37.954084Z digest=sha256:2ddfbf7439c3704bd0d8465a07c5d533c46ee3f7169009a50209135d999c12d0

Observation aa57d38e-4bff-46db-a5f4-b09d9ad51187 · inbound

Doomed from the Start: Early Abort of LLM Agent Episodes via a Recall-Controlled Probe Cascade cites this paper.

Doomed from the Start: Early Abort of LLM Agent Episodes via a Recall-Controlled Probe Cascade AgentGym: Evolving Large Language Model-based Agents across Diverse Environments

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-02T08:18:06.225163Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T08:18:06.225163Z digest=sha256:a3699a15f88dc88c6526cb472264654b98d020267e71c5f7fd8a8ebae6417831

Observation 84e3e9a8-6a83-4214-b5d6-bfda514d015f · inbound

Evolutionary Intelligence for Scientific Discovery: From Evolutionary Computation to Cumulative Discovery Systems cites this paper.

Evolutionary Intelligence for Scientific Discovery: From Evolutionary Computation to Cumulative Discovery Systems AgentGym: Evolving Large Language Model-based Agents across Diverse Environments

Reference 41

Resolution
unresolved
no resolver link, observed 2026-07-13T00:55:17.107245Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-13T00:55:17.107245Z digest=sha256:dea78e383b8a722c6131122dc2ceefe1fc4d2a8c54c1ab8d41cad1dc27dda412

Observation dce7f585-c72a-4faf-bd27-b9cb0d912ec5 · inbound

From Trajectories to Prefixes: Reusing Teacher Trajectories via Replayed Prefixes and Online Continuation cites this paper.

From Trajectories to Prefixes: Reusing Teacher Trajectories via Replayed Prefixes and Online Continuation AgentGym: Evolving Large Language Model-based Agents across Diverse Environments

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-02T08:56:35.622302Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T08:56:35.622302Z digest=sha256:c63a09a0c25e0fd9da292ec630c3f1d4b5d79474b7afd828cfd145342b344e67

Observation e95b5aef-30eb-4297-9537-ef34467951ea · inbound

Agent Security Needs Redefinition through a Holistic Framework cites this paper.

Agent Security Needs Redefinition through a Holistic Framework AgentGym: Evolving Large Language Model-based Agents across Diverse Environments

Reference 218

Resolution
unresolved
no resolver link, observed 2026-08-01T06:04:46.361666Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T06:04:46.361666Z digest=sha256:66826abd6edb49c4fb01f083220887fe8e384d70bf4cf10772879dd3e61f1473

Observation 8eacd945-76f1-463c-93e5-fdfb59c3d38c · inbound

AgentOmnia: Scaling Agentic Models for Full-Scenario Applications cites this paper.

AgentOmnia: Scaling Agentic Models for Full-Scenario Applications AgentGym: Evolving Large Language Model-based Agents across Diverse Environments

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-01T03:38:18.754752Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T03:38:18.754752Z digest=sha256:2554c2a472e1a064768b4a09d82dd98312dd3234ba32a8e63031e193cb2b540f

Observation 8eedf16f-c33f-4993-a960-575a7664928a · inbound

NeSyFS: A Neuro-symbolic Fast-Slow Thinking Framework for LLM Agent under Partial Observability cites this paper.

NeSyFS: A Neuro-symbolic Fast-Slow Thinking Framework for LLM Agent under Partial Observability AgentGym: Evolving Large Language Model-based Agents across Diverse Environments

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-03T16:54:58.749336Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-03T16:54:58.749336Z digest=sha256:14b2b385ce8ab417e2cea9522e7f57ef4b88188f455d77d539da41625e3c98bf

Observation 21508525-cb10-4213-94b5-b0911f56fe6b · inbound

Instruction-Conditioned Exploration with Asymmetric Reinforcement Learning and Self-Distillation cites this paper.

Instruction-Conditioned Exploration with Asymmetric Reinforcement Learning and Self-Distillation AgentGym: Evolving Large Language Model-based Agents across Diverse Environments

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-04T15:25:49.227255Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T15:25:49.227255Z digest=sha256:c7aefa18c555fe6b0f83b0c0ca108f6b9205791db725c86d2603061e84037bc0