Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 7 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 4 inbound Pith citation observations for arXiv:2202.04628.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-07T14:56:24.865596Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-05-16T06:32:27.300616Z
0 of 0 outbound references displayed
External citation measurements
No source-named external measurement is stored.
No outbound reference observations are available for this paper version.
Observation 331db90b-2b4a-4eca-a020-5fd6092ffc97 · inbound
Towards VM Rescheduling Optimization Through Deep Reinforcement Learning Reinforcement Learning with Sparse Rewards using Guidance from Offline Demonstration
Reference 54
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 005feb40-46b3-4527-97ce-533f5247ca9b · inbound
LaViPlan : Language-Guided Visual Path Planning with RLVR Reinforcement Learning with Sparse Rewards using Guidance from Offline Demonstration
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4387cb75-0f5b-465d-b19b-f235d378633f · inbound
AceGRPO: Adaptive Curriculum Enhanced Group Relative Policy Optimization for Autonomous Machine Learning Engineering Reinforcement Learning with Sparse Rewards using Guidance from Offline Demonstration
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation dba72cb6-62df-4dd0-96f0-98af582b032c · inbound
Expert Behavior Prior Reinforcement Learning Reinforcement Learning with Sparse Rewards using Guidance from Offline Demonstration
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.