Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 7 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 11 inbound Pith citation observations for arXiv:2110.02034.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-07T00:32:51.089069Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-05-22T23:22:15.583985Z
0 of 0 outbound references displayed
External citation measurements
No source-named external measurement is stored.
No outbound reference observations are available for this paper version.
Observation 2514533e-c7a8-473d-b31b-8ec9b4360ba1 · inbound
Learning to Play Piano in the Real World Dropout Q-Functions for Doubly Efficient Reinforcement Learning
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 962346cb-77aa-4e31-a6e4-c439b653c794 · inbound
The Courage to Stop: Overcoming Sunk Cost Fallacy in Deep Reinforcement Learning Dropout Q-Functions for Doubly Efficient Reinforcement Learning
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a378b549-95f8-4ce0-962c-75f61d17b505 · inbound
Squeeze the Soaked Sponge: Efficient Off-policy Reinforcement Finetuning for Large Language Model Dropout Q-Functions for Doubly Efficient Reinforcement Learning
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 740cbb18-564c-415a-b479-54626d0a7b54 · inbound
Reinforcement learning entangling operations on spin qubits Dropout Q-Functions for Doubly Efficient Reinforcement Learning
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5d3bcf8a-fce9-414f-9150-1a6f27f5b1a5 · inbound
Low-Rank Adaptation for Critic Learning in Off-Policy Reinforcement Learning Dropout Q-Functions for Doubly Efficient Reinforcement Learning
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation a8075e65-fb29-4fc6-b12f-be8481159250 · inbound
Distributional Value Estimation Without Target Networks for Robust Quality-Diversity Dropout Q-Functions for Doubly Efficient Reinforcement Learning
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 832ce788-dbce-43e4-95c0-4292be79cd47 · inbound
SOPE: Stabilizing Off-Policy Evaluation for Online RL with Prior Data Dropout Q-Functions for Doubly Efficient Reinforcement Learning
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 5a677ca3-750c-4b1a-93cd-c9adc57f5421 · inbound
SOPE: Stabilizing Off-Policy Evaluation for Online RL with Prior Data Dropout Q-Functions for Doubly Efficient Reinforcement Learning
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 1adb0229-71e1-45d3-b114-f389dcb864d4 · inbound
Debiased Model-based Representations for Sample-efficient Continuous Control Dropout Q-Functions for Doubly Efficient Reinforcement Learning
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation d8ce6f9a-7e3d-42fd-89cc-5cd12670f129 · inbound
Deep Reinforcement Learning: From First Principles to Reasoning Models Dropout Q-Functions for Doubly Efficient Reinforcement Learning
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 18432877-8ad2-4789-bcc2-c2397eaece94 · inbound
ReBRAC-v2: The Return of the King Dropout Q-Functions for Doubly Efficient Reinforcement Learning
Reference 73
Source-reported events for the cited work
Unavailable: canonical work link unavailable.