Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T05:37:46.402936Z
Paper Citation Record · LEDGER
As of 7 August 2026, this Paper Citation Record lists 23 of 23 outbound references and 8 inbound Pith citation observations for arXiv:2506.07505.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T05:37:46.402936Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-05T04:27:43.858412Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-05-19T05:12:05.250025Z
23 of 23 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 54098eb7-f739-4b53-b9d7-90eaa0d274cc · outbound
Reinforcement Learning via Implicit Imitation Guidance Efficient Online Reinforcement Learning with Offline Data
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c0c7ca42-7fe9-4e31-b0ba-49f239dff920 · outbound
Reinforcement Learning via Implicit Imitation Guidance epochs per update
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation fe3a2452-6743-4417-a163-f222b9a72661 · outbound
Reinforcement Learning via Implicit Imitation Guidance worse", which are successful demonstrations collected by inexperienced operators to incorporate additional diversity. Note that even though it is labeled
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 532c19c2-6b0f-4037-997e-8780c6c955fc · outbound
Reinforcement Learning via Implicit Imitation Guidance Imitation Bootstrapped Reinforcement Learning
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 06c7343a-0aec-4818-8248-c1f59cbebcfd · outbound
Reinforcement Learning via Implicit Imitation Guidance Offline Reinforcement Learning with Implicit Q-Learning
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b2329927-ac49-4b9c-b574-274bd0fd61ed · outbound
Reinforcement Learning via Implicit Imitation Guidance Offline Retraining for Online RL: Decoupled Policy Learning to Mitigate Exploration Bias
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a13b6330-ea3a-4e6e-8793-87762ccd29da · outbound
Reinforcement Learning via Implicit Imitation Guidance Over- coming exploration in reinforcement learning with demonstrations
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 2d76a449-f55e-45b6-9a25-4fe50f74881b · outbound
Reinforcement Learning via Implicit Imitation Guidance Computational Theories of Curiosity-Driven Learning
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation aeaa12dd-e2cd-41c4-8703-41c978fb7710 · outbound
Reinforcement Learning via Implicit Imitation Guidance Learning Complex Dexterous Manipulation with Deep Reinforcement Learning and Demonstrations
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8a60e0dd-1666-43bd-8c84-0e5cc8d313af · outbound
Reinforcement Learning via Implicit Imitation Guidance Schmidhuber
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 13036d68-8079-4468-a573-ac508a627241 · outbound
Reinforcement Learning via Implicit Imitation Guidance Parrot: Data-Driven Behavioral Priors for Reinforcement Learning
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0b408ba0-19b1-410c-bd43-e0e121d4bbbf · outbound
Reinforcement Learning via Implicit Imitation Guidance Hybrid RL: Using Both Offline and Online Data Can Make RL Efficient
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4474229d-0083-4b59-87ee-9d01368b8b76 · outbound
Reinforcement Learning via Implicit Imitation Guidance Leveraging Demonstrations for Deep Reinforcement Learning on Robotics Problems with Sparse Rewards
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0066d2d0-370f-462f-b277-9688d2779557 · outbound
Reinforcement Learning via Implicit Imitation Guidance Learning latent state representation for speeding up exploration
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 5da8b341-7bea-4905-9769-d2ff68f8c0dd · outbound
Reinforcement Learning via Implicit Imitation Guidance Policy Expansion for Bridging Offline-to-Online Reinforcement Learning
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 14ab7d25-fb58-442a-b0b2-d46ed3aa6d56 · outbound
Reinforcement Learning via Implicit Imitation Guidance Exploration by Random Network Distillation
Reference 2016
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fd8c01a0-028f-4001-ae53-4883703f0cc9 · outbound
Reinforcement Learning via Implicit Imitation Guidance Learning by Playing - Solving Sparse Reward Tasks from Scratch
Reference 2017
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 4197caaf-9d9d-41ea-b2dc-43b242038ebc · outbound
Reinforcement Learning via Implicit Imitation Guidance AWAC: Accelerating Online Reinforcement Learning with Offline Datasets
Reference 2018
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 084fe3ea-9951-45c1-9f97-44b59aa7afea · outbound
Reinforcement Learning via Implicit Imitation Guidance Self-Supervised Exploration via Disagreement
Reference 2019
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 18e26509-53b6-4ac0-a579-ec7add5f2be0 · outbound
Reinforcement Learning via Implicit Imitation Guidance Go-Explore: a New Approach for Hard-Exploration Problems
Reference 2020
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 59835236-da50-444b-87cb-09f63dc28f21 · outbound
Reinforcement Learning via Implicit Imitation Guidance Efficient Exploration via State Marginal Matching
Reference 2021
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 88953db2-4e1a-4aa0-8f4b-bd80c719fd01 · outbound
Reinforcement Learning via Implicit Imitation Guidance Making Efficient Use of Demonstrations to Solve Hard Exploration Problems
Reference 2022
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 88e13b0c-c32f-4da1-9331-d149b31a9721 · outbound
Reinforcement Learning via Implicit Imitation Guidance MoDem: Accelerating Visual Model-Based Reinforcement Learning with Demonstrations
Reference 2023
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f7f7acb0-042e-4fab-b930-ab01711ff06d · inbound
EXPO: Stable Reinforcement Learning with Expressive Policies Reinforcement Learning via Implicit Imitation Guidance
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 912f030f-4650-4d47-b137-548db7d98738 · inbound
Value Flows Reinforcement Learning via Implicit Imitation Guidance
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 73bb73d3-8468-4d6c-b222-9bd2e736b647 · inbound
Fine-tuning is Not Enough: A Parallel Framework for Collaborative Imitation and Reinforcement Learning in End-to-end Autonomous Driving Reinforcement Learning via Implicit Imitation Guidance
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 740ec016-c10b-4d8c-b869-42282c0b9244 · inbound
Incremental Residual Reinforcement Learning Toward Real-World Learning for Social Navigation Reinforcement Learning via Implicit Imitation Guidance
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 510bd843-417c-4782-93d9-602cb0adf3fb · inbound
Incremental Residual Reinforcement Learning Toward Real-World Learning for Social Navigation Reinforcement Learning via Implicit Imitation Guidance
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2eee268b-bfe8-4404-b5b0-805b788a900a · inbound
FASTER: Value-Guided Sampling for Fast RL Reinforcement Learning via Implicit Imitation Guidance
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 6dc67857-44d7-4bbf-85f2-b9e352933bc6 · inbound
Do You Really Need to Pretrain Q-Functions for Online RL Fine-Tuning? Reinforcement Learning via Implicit Imitation Guidance
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 49b28c9e-8205-4ea0-b167-2c60e716a353 · inbound
Do You Really Need to Pretrain Q-Functions for Online RL Fine-Tuning? Reinforcement Learning via Implicit Imitation Guidance
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.