Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 7 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 57 inbound Pith citation observations for arXiv:2411.02337.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-07T14:37:46.329762Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-07-04T17:20:00.041093Z
0 of 0 outbound references displayed
External citation measurements
No source-named external measurement is stored.
No outbound reference observations are available for this paper version.
Observation 1a439e27-a856-4426-b5ad-49f735016f73 · inbound
Plan-and-Act: Improving Planning of Agents for Long-Horizon Tasks WebRL: Training LLM Web Agents via Self-Evolving Online Curriculum Reinforcement Learning
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 8a08d509-760c-4940-bbdb-8350116a3a5f · inbound
Advances and Challenges in Foundation Agents: From Brain-Inspired Intelligence to Evolutionary, Collaborative, and Safe Systems WebRL: Training LLM Web Agents via Self-Evolving Online Curriculum Reinforcement Learning
Reference 143
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 4efd62e6-4159-4d0d-b049-b694982913da · inbound
ProgRM: Build Better GUI Agents with Progress Rewards WebRL: Training LLM Web Agents via Self-Evolving Online Curriculum Reinforcement Learning
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c754b48f-8e54-47d5-923c-ac41bbb93d9f · inbound
Large Language Models for Planning: A Comprehensive and Systematic Survey WebRL: Training LLM Web Agents via Self-Evolving Online Curriculum Reinforcement Learning
Reference 195
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 84a7a678-435d-4764-9a4c-23ce173b09bc · inbound
Agent-Environment Alignment via Automated Interface Generation WebRL: Training LLM Web Agents via Self-Evolving Online Curriculum Reinforcement Learning
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f94e7756-8544-49e9-941d-b4d6cd891db6 · inbound
AgentDNS: A Root Domain Naming System for LLM Agents WebRL: Training LLM Web Agents via Self-Evolving Online Curriculum Reinforcement Learning
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 767db96d-c03f-4295-8692-071b489e8168 · inbound
ZeroGUI: Automating Online GUI Learning at Zero Human Cost WebRL: Training LLM Web Agents via Self-Evolving Online Curriculum Reinforcement Learning
Reference 47
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 14a271ba-3cbd-410d-abf3-89156620648a · inbound
OWL: Optimized Workforce Learning for General Multi-Agent Assistance in Real-World Task Automation WebRL: Training LLM Web Agents via Self-Evolving Online Curriculum Reinforcement Learning
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation dcba1132-d734-494e-be0a-fc5a787f1ed5 · inbound
Self-Challenging Language Model Agents WebRL: Training LLM Web Agents via Self-Evolving Online Curriculum Reinforcement Learning
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 59ccd5d8-9b02-48c3-8bd2-ba65dad8f4ba · inbound
Orak: A Foundational Benchmark for Training and Evaluating LLM Agents on Diverse Video Games WebRL: Training LLM Web Agents via Self-Evolving Online Curriculum Reinforcement Learning
Reference 51
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 7498d1e4-28c7-4c4e-bd30-e41f5108b9a1 · inbound
Truly Self-Improving Agents Require Intrinsic Metacognitive Learning WebRL: Training LLM Web Agents via Self-Evolving Online Curriculum Reinforcement Learning
Reference 57
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2e782046-7782-462f-9b1e-06e866525791 · inbound
Thinking vs. Doing: Agents that Reason by Scaling Test-Time Interaction WebRL: Training LLM Web Agents via Self-Evolving Online Curriculum Reinforcement Learning
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 23b686ae-f595-477c-a155-9ea064784556 · inbound
Atomic-to-Compositional Generalization for Mobile Agents with A New Benchmark and Scheduling System WebRL: Training LLM Web Agents via Self-Evolving Online Curriculum Reinforcement Learning
Reference 41
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 584b1166-6f2e-4fa4-b163-63d343223acf · inbound
Build the web for agents, not agents for the web WebRL: Training LLM Web Agents via Self-Evolving Online Curriculum Reinforcement Learning
Reference 46
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e6aefe98-2f29-4ad0-b8e9-c6a42f5efe51 · inbound
Agent-RLVR: Training Software Engineering Agents via Guidance and Environment Rewards WebRL: Training LLM Web Agents via Self-Evolving Online Curriculum Reinforcement Learning
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation eea4fae6-f494-4e05-9c8d-05c761871989 · inbound
MEM1: Learning to Synergize Memory and Reasoning for Efficient Long-Horizon Agents WebRL: Training LLM Web Agents via Self-Evolving Online Curriculum Reinforcement Learning
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 4a1f1581-f669-4224-9f39-26e04b903bbc · inbound
WebSynthesis: World-Model-Guided MCTS for Efficient WebUI-Trajectory Synthesis WebRL: Training LLM Web Agents via Self-Evolving Online Curriculum Reinforcement Learning
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d5f4170c-1008-4f8d-881f-50d4d395223b · inbound
A Survey of Self-Evolving Agents: What, When, How, and Where to Evolve on the Path to Artificial Super Intelligence WebRL: Training LLM Web Agents via Self-Evolving Online Curriculum Reinforcement Learning
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 56fc05ce-ba75-47ee-88b3-2906e94b33b0 · inbound
SEAgent: Self-Evolving Computer Use Agent with Autonomous Learning from Experience WebRL: Training LLM Web Agents via Self-Evolving Online Curriculum Reinforcement Learning
Reference 47
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3420f575-8354-49a0-a0ac-002bd0281248 · inbound
Cognitive Duality for Adaptive Web Agents WebRL: Training LLM Web Agents via Self-Evolving Online Curriculum Reinforcement Learning
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8ebe02db-d3f0-423e-8c0f-2cf95809e21e · inbound
EvoCurr: Self-evolving Curriculum with Behavior Code Generation for Complex Decision-making WebRL: Training LLM Web Agents via Self-Evolving Online Curriculum Reinforcement Learning
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a32608b8-1845-430f-b8b0-6d33fa8819dd · inbound
Atom-Searcher: Enhancing Agentic Deep Research via Fine-Grained Atomic Thought Reward WebRL: Training LLM Web Agents via Self-Evolving Online Curriculum Reinforcement Learning
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ceb3ec28-6bb5-463b-97f4-03a7fcc6ba6e · inbound
CODA: Coordinating the Cerebrum and Cerebellum for a Dual-Brain Computer Use Agent with Decoupled Reinforcement Learning WebRL: Training LLM Web Agents via Self-Evolving Online Curriculum Reinforcement Learning
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6841f209-0436-4463-aefc-e54ffc9774f9 · inbound
Symbolic Graphics Programming with Large Language Models WebRL: Training LLM Web Agents via Self-Evolving Online Curriculum Reinforcement Learning
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 404606f8-a388-4e91-bd77-0ac751fe4792 · inbound
A global log for medical AI WebRL: Training LLM Web Agents via Self-Evolving Online Curriculum Reinforcement Learning
Reference 123
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 456f2f7e-ae95-44e5-9e90-486f0853b895 · inbound
Agent Learning via Early Experience WebRL: Training LLM Web Agents via Self-Evolving Online Curriculum Reinforcement Learning
Reference 43
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3dd41f15-24e5-4f99-a39f-ff294db464cd · inbound
From Refusal to Recovery: A Control-Theoretic Approach to Generative AI Guardrails WebRL: Training LLM Web Agents via Self-Evolving Online Curriculum Reinforcement Learning
Reference 62
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation d6916505-eaec-47c5-8552-e8d1b831704b · inbound
IPR-1: Interactive Physical Reasoner WebRL: Training LLM Web Agents via Self-Evolving Online Curriculum Reinforcement Learning
Reference 48
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 52a3c39b-62df-47b7-9026-a1735bc6fb0e · inbound
IPR-1: Interactive Physical Reasoner WebRL: Training LLM Web Agents via Self-Evolving Online Curriculum Reinforcement Learning
Reference 48
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 59479678-5c43-4a1e-b726-d3dfcd61f296 · inbound
DynaWeb: Model-Based Reinforcement Learning of Web Agents WebRL: Training LLM Web Agents via Self-Evolving Online Curriculum Reinforcement Learning
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 577a560a-e36c-4b7c-91b0-9fe6805dd2df · inbound
GUI-Libra: Training Native GUI Agents to Reason and Act with Action-aware Supervision and Partially Verifiable RL WebRL: Training LLM Web Agents via Self-Evolving Online Curriculum Reinforcement Learning
Reference 45
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8b693a41-6a3b-4de6-9082-f74d0f1123f6 · inbound
From Pixels to Digital Agents: An Empirical Study on the Taxonomy and Technological Trends of Reinforcement Learning Environments WebRL: Training LLM Web Agents via Self-Evolving Online Curriculum Reinforcement Learning
Reference 208
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 1b0603ab-9dfe-4f64-bc5d-bb4f5a41243e · inbound
What's Missing in Screen-to-Action? Towards a UI-in-the-Loop Paradigm for Multimodal GUI Reasoning WebRL: Training LLM Web Agents via Self-Evolving Online Curriculum Reinforcement Learning
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 79d01310-4cf2-4dc6-a202-fc8fec2f88f1 · inbound
AIT Academy: Cultivating the Complete Agent with a Confucian Three-Domain Curriculum WebRL: Training LLM Web Agents via Self-Evolving Online Curriculum Reinforcement Learning
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 666a29e3-17d5-4c68-ad9e-ed837f4bdb64 · inbound
Improving LLM Code Generation via Requirement-Aware Curriculum Reinforcement Learning WebRL: Training LLM Web Agents via Self-Evolving Online Curriculum Reinforcement Learning
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation f2208569-4677-4a3c-bb48-212c0babdade · inbound
Generate, Filter, Control, Replay: A Comprehensive Survey of Rollout Strategies for LLM Reinforcement Learning WebRL: Training LLM Web Agents via Self-Evolving Online Curriculum Reinforcement Learning
Reference 92
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation e0674fe6-1e47-49fd-85ed-65a97b37eda1 · inbound
Milestone-Guided Policy Learning for Long-Horizon Language Agents WebRL: Training LLM Web Agents via Self-Evolving Online Curriculum Reinforcement Learning
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 058e9db2-225c-470b-902c-fa57a3760c99 · inbound
Weblica: Scalable and Reproducible Training Environments for Visual Web Agents WebRL: Training LLM Web Agents via Self-Evolving Online Curriculum Reinforcement Learning
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 11ce60e9-4099-45ee-8ddb-43cd23bb6da5 · inbound
SOD: Step-wise On-policy Distillation for Small Language Model Agents WebRL: Training LLM Web Agents via Self-Evolving Online Curriculum Reinforcement Learning
Reference 49
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 61a65632-bb1d-4a68-a58d-f995578abf69 · inbound
SOD: Step-wise On-policy Distillation for Small Language Model Agents WebRL: Training LLM Web Agents via Self-Evolving Online Curriculum Reinforcement Learning
Reference 49
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0ed81619-69fc-4fd7-9a72-49183809b3a4 · inbound
SimWorld Studio: Automatic Environment Generation with Evolving Coding Agent for Embodied Agent Learning WebRL: Training LLM Web Agents via Self-Evolving Online Curriculum Reinforcement Learning
Reference 62
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation bd028f99-e4e3-48ac-b386-d6b57c581225 · inbound
SimWorld Studio: Automatic Environment Generation with Evolving Coding Agent for Embodied Agent Learning WebRL: Training LLM Web Agents via Self-Evolving Online Curriculum Reinforcement Learning
Reference 62
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation b069b3c8-cb86-418f-94b9-e3117a4981e0 · inbound
Rewarding Beliefs, Not Actions: Consistency-Guided Credit Assignment for Long-Horizon Agents WebRL: Training LLM Web Agents via Self-Evolving Online Curriculum Reinforcement Learning
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation ed399df6-0832-4121-919d-a59193c3a5bc · inbound
Weasel: Out-of-Domain Generalization for Web Agents via Importance-Diversity Data Selection WebRL: Training LLM Web Agents via Self-Evolving Online Curriculum Reinforcement Learning
Reference 48
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation c62bc5f8-cedb-4ae3-a1b4-4bcd212971a4 · inbound
Weasel: Out-of-Domain Generalization for Web Agents via Importance-Diversity Data Selection WebRL: Training LLM Web Agents via Self-Evolving Online Curriculum Reinforcement Learning
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 6e5bcc9a-5671-4435-ad9f-e72c017169f0 · inbound
Mem-$\pi$: Adaptive Memory through Learning When and What to Generate WebRL: Training LLM Web Agents via Self-Evolving Online Curriculum Reinforcement Learning
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 6557edbf-7468-470d-a8f4-45441a6db87a · inbound
DRIVE: Modeling Skills at the Reasoning and Interaction Levels for Web Agents under Continual Learning WebRL: Training LLM Web Agents via Self-Evolving Online Curriculum Reinforcement Learning
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 3d93c6be-e952-4002-9d0c-81f60346f689 · inbound
Learning to Act under Noise: Enhancing Agent Robustness via Noisy Environments WebRL: Training LLM Web Agents via Self-Evolving Online Curriculum Reinforcement Learning
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 9821d9a6-fad3-4048-a98a-743e8d7f257d · inbound
Deep Research as Rubric for Reinforcement Learning WebRL: Training LLM Web Agents via Self-Evolving Online Curriculum Reinforcement Learning
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation abd50648-c9f2-4285-9a0d-9ff79f69f7f6 · inbound
AliyunConsoleAgent: Training Web Agents in Real-World Cloud Environments via Distillation and Reinforcement Learning WebRL: Training LLM Web Agents via Self-Evolving Online Curriculum Reinforcement Learning
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 9bcde929-0217-494a-ba96-38d72605a508 · inbound
Speculative Rollback Correction for Quality-Diverse Web Agent Imitation WebRL: Training LLM Web Agents via Self-Evolving Online Curriculum Reinforcement Learning
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 0ce9644d-fea6-4905-a03a-468dc6aaa1ec · inbound
Training the Orchestrator: A Supervised Approach to End-to-End PDDL Planning with LLM Agents WebRL: Training LLM Web Agents via Self-Evolving Online Curriculum Reinforcement Learning
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 8abca0e9-e8ab-472b-b5fc-d427f2488334 · inbound
Escaping the Self-Confirmation Trap: An Execute-Distill-Verify Paradigm for Agentic Experience Learning WebRL: Training LLM Web Agents via Self-Evolving Online Curriculum Reinforcement Learning
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 18440b04-b9a5-4dea-b32d-f6f4e07bcc55 · inbound
Agentic-DPO: From Imitation to Agentic Policy Optimization on Expert Trajectories WebRL: Training LLM Web Agents via Self-Evolving Online Curriculum Reinforcement Learning
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3ea66aef-dc4e-4898-8121-a7bf36a89eb6 · inbound
SCALECUA: Scaling Computer Use Agents with Verifiable Task Synthesis and Efficient Online RL WebRL: Training LLM Web Agents via Self-Evolving Online Curriculum Reinforcement Learning
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 258eb65e-660f-489a-ac9d-68331475c65b · inbound
OSReward: Instituting Standardized Evaluation for Cross-Platform Computer-Use Reward Models WebRL: Training LLM Web Agents via Self-Evolving Online Curriculum Reinforcement Learning
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2ea62fda-50e5-4f86-b47c-1711cdedaf18 · inbound
Progressive Agent Skill Generation via Reinforcement Learning WebRL: Training LLM Web Agents via Self-Evolving Online Curriculum Reinforcement Learning
Reference 74
Source-reported events for the cited work
Unavailable: canonical work link unavailable.