Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 8 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 19 inbound Pith citation observations for arXiv:2302.02948.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-07T14:14:45.188595Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-07-04T16:49:57.181580Z
0 of 0 outbound references displayed
External citation measurements
No source-named external measurement is stored.
No outbound reference observations are available for this paper version.
Observation 07b8cd2f-ab37-4319-9b16-b246748fc2b7 · inbound
IDQL: Implicit Q-Learning as an Actor-Critic Method with Diffusion Policies Efficient Online Reinforcement Learning with Offline Data
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation da1fdad5-7c8d-4786-9ebf-e5c3fdaf1061 · inbound
Reinforcement Learning with Foundation Priors: Let the Embodied Agent Efficiently Learn on Its Own Efficient Online Reinforcement Learning with Offline Data
Reference 64
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 9ca01919-e7c7-4a62-9b55-695b0f199b9b · inbound
Learning to Trust Bellman Updates: Selective State-Adaptive Regularization for Offline RL Efficient Online Reinforcement Learning with Offline Data
Reference 2022
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 54098eb7-f739-4b53-b9d7-90eaa0d274cc · inbound
Reinforcement Learning via Implicit Imitation Guidance Efficient Online Reinforcement Learning with Offline Data
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 19b8f177-1080-4532-812f-1a5351fb1d6d · inbound
Decentralized Relaxed Smooth Optimization with Gradient Descent Methods Efficient Online Reinforcement Learning with Offline Data
Reference 2022
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6af8edd1-46b1-4419-882f-c5d3db83d8a5 · inbound
Value Flows Efficient Online Reinforcement Learning with Offline Data
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8ea21e37-3b0e-49ab-9fd5-8a67e61b2199 · inbound
Self-Supervised Multisensory Pretraining for Contact-Rich Robot Reinforcement Learning Efficient Online Reinforcement Learning with Offline Data
Reference 44
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 7a7d4669-fe08-4290-ab25-458ba3ca73e6 · inbound
Self-Supervised Multisensory Pretraining for Contact-Rich Robot Reinforcement Learning Efficient Online Reinforcement Learning with Offline Data
Reference 44
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 93d52c32-20b8-454a-bc34-a47bfcea627a · inbound
Online World Modeling Enables Real-World Inverse Reinforcement Learning from Observation Efficient Online Reinforcement Learning with Offline Data
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 44b8beff-9de2-4cae-95ce-d09ce9fe70b4 · inbound
Long-Horizon Q-Learning: Accurate Value Learning via n-Step Inequalities Efficient Online Reinforcement Learning with Offline Data
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation f2bb1677-e9ad-4059-94aa-336f814d5aae · inbound
Long-Horizon Q-Learning: Accurate Value Learning via n-Step Inequalities Efficient Online Reinforcement Learning with Offline Data
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation e4ef87bb-5bf0-42bd-9790-d38f19d2060e · inbound
RankQ: Offline-to-Online Reinforcement Learning via Self-Supervised Action Ranking Efficient Online Reinforcement Learning with Offline Data
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 4302685f-03af-401c-bfd5-d9e21e29b59e · inbound
RankQ: Offline-to-Online Reinforcement Learning via Self-Supervised Action Ranking Efficient Online Reinforcement Learning with Offline Data
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation d19e3344-8788-4200-ae8b-6b7fb000eb31 · inbound
Improving Robotic Generalist Policies via Flow Reversal Steering Efficient Online Reinforcement Learning with Offline Data
Reference 92
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 3122a81d-378e-485c-97e4-c1ebfff421a5 · inbound
Learning Process Rewards via Success Visitation Matching for Efficient RL Efficient Online Reinforcement Learning with Offline Data
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 3bf457a2-ff38-4369-8ac5-5166c4c1a1ca · inbound
An Introduction to Causal Reinforcement Learning Efficient Online Reinforcement Learning with Offline Data
Reference 140
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 63993894-2958-492b-a60f-a338fef6ccec · inbound
OmniTacTune: Policy-Agnostic Real-World RL for Tactile Residual Adaptation of Visual Policies Efficient Online Reinforcement Learning with Offline Data
Reference 58
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b0587974-934a-47ec-a15b-e711c698d88d · inbound
Do You Really Need to Pretrain Q-Functions for Online RL Fine-Tuning? Efficient Online Reinforcement Learning with Offline Data
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fc8d6680-afd6-4aa9-abff-3fbe039a3cd8 · inbound
Do You Really Need to Pretrain Q-Functions for Online RL Fine-Tuning? Efficient Online Reinforcement Learning with Offline Data
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.