Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 7 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 12 inbound Pith citation observations for arXiv:2406.11176.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-07T13:51:42.816272Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-07-04T13:59:51.871553Z
0 of 0 outbound references displayed
External citation measurements
No source-named external measurement is stored.
No outbound reference observations are available for this paper version.
Observation 9eb5afc0-64c0-48da-b24b-8ccaca91f770 · inbound
RRO: LLM Agent Optimization Through Rising Reward Trajectories Watch Every Step! LLM Agent Learning via Iterative Step-Level Process Refinement
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b0282fc0-9feb-4ac3-a30a-0f6c441ec3e6 · inbound
ARIA: Training Language Agents with Intention-Driven Reward Aggregation Watch Every Step! LLM Agent Learning via Iterative Step-Level Process Refinement
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b1f80aea-97f8-48a6-95d1-188bcd44a41f · inbound
LLaPipe: LLM-Guided Reinforcement Learning for Automated Data Preparation Pipeline Construction Watch Every Step! LLM Agent Learning via Iterative Step-Level Process Refinement
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d41d0aa5-2fdd-42f2-b98b-92a9fc215029 · inbound
Source Component Shift Adaptation via Offline Decomposition and Online Mixing Approach Watch Every Step! LLM Agent Learning via Iterative Step-Level Process Refinement
Reference 63
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b8a50294-3d51-4ab8-b6ae-ac7158af8ec7 · inbound
Leveraging OS-Level Primitives for Robotic Action Management Watch Every Step! LLM Agent Learning via Iterative Step-Level Process Refinement
Reference 69
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e141f2e5-c99d-43d4-acb4-8e21cc1a7426 · inbound
MENTOR: Reinforcement Learning via Flexible Teacher-Optimized Rewards for Tool-Use Distillation Watch Every Step! LLM Agent Learning via Iterative Step-Level Process Refinement
Reference 56
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d03d49af-80a8-499d-b8ca-7a88a1d5fd69 · inbound
From Coarse to Fine: Self-Adaptive Hierarchical Planning for LLM Agents Watch Every Step! LLM Agent Learning via Iterative Step-Level Process Refinement
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 17c7d3ff-d50a-45d5-868b-bfe81a49f6fa · inbound
Rewarding Beliefs, Not Actions: Consistency-Guided Credit Assignment for Long-Horizon Agents Watch Every Step! LLM Agent Learning via Iterative Step-Level Process Refinement
Reference 40
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation e7d12215-7382-44ad-8576-043a4b10013d · inbound
MetaPS: Adaptive Programmatic Strategy Selection for Market Agents Watch Every Step! LLM Agent Learning via Iterative Step-Level Process Refinement
Reference 74
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 25f690f8-8941-470c-999a-6c17f4848234 · inbound
Reinforcement Learning without Ground-Truth Solutions can Improve LLMs Watch Every Step! LLM Agent Learning via Iterative Step-Level Process Refinement
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation b00c4d8e-0e9d-4b43-a145-7500acd36a99 · inbound
Fishing Out Free Riders: Shapley-Based Reward Attribution for Parallel Reasoning via Reinforcement Learning Watch Every Step! LLM Agent Learning via Iterative Step-Level Process Refinement
Reference 57
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4c53a9ff-d3fa-4631-bbf8-a4207098375a · inbound
Leveraging Trajectory Graphs for Pre-Execution Error Diagnosis in Agentic LLM Systems Watch Every Step! LLM Agent Learning via Iterative Step-Level Process Refinement
Reference 64
Source-reported events for the cited work
Unavailable: canonical work link unavailable.