Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-10T21:32:16.181011Z
Paper Citation Record · LEDGER
As of 11 August 2026, this Paper Citation Record lists 20 of 20 outbound references and 6 inbound Pith citation observations for arXiv:2501.04870.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-10T21:32:16.181011Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-11T06:34:44.6726+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-06T21:19:10.654724Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-05-20T23:49:14.996703Z
20 of 20 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation d0ec2c96-cbe6-47d6-965c-8c40046abc08 · outbound
Deep Transfer $Q$-Learning for Offline Non-Stationary Reinforcement Learning Offline Multi-task Transfer RL with Representational Penalization
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2caaed7d-f951-4f6e-a72a-c1d286d793d3 · outbound
Deep Transfer $Q$-Learning for Offline Non-Stationary Reinforcement Learning After some algebra we get that ∥ bQp t − Q∗ agg t ∥2 nM,bPagg t ≤ ∥gp t − Q∗ agg t ∥2 nM,bPagg t + 2 nM nMX i=1 (byrwt−ki t,i − Q∗ agg t,i ) · ( bQp t,i − gp t,i)
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation fcc1872c-9d52-44da-8afe-7255ce3a95a5 · outbound
Deep Transfer $Q$-Learning for Offline Non-Stationary Reinforcement Learning Unresolved cited work
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 7575b9fe-bbf0-4a72-bd2f-5b7d84f48565 · outbound
Deep Transfer $Q$-Learning for Offline Non-Stationary Reinforcement Learning sub-Gaussian random variables with variance parameter σ
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 14068bac-a517-45fd-885d-c40053615633 · outbound
Deep Transfer $Q$-Learning for Offline Non-Stationary Reinforcement Learning copies of z, G be a b-uniformly-bounded function class satisfying log(N∞(ϵ, G, zn 1 )) ≤ v log ebn ϵ for some quantity v
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 4c3e234b-a9cc-4f49-8492-0e8e88f5e2a4 · outbound
Deep Transfer $Q$-Learning for Offline Non-Stationary Reinforcement Learning Robust angle-based transfer learning in high dimensions
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 322b8543-e945-4058-bdad-7302491fa76e · outbound
Deep Transfer $Q$-Learning for Offline Non-Stationary Reinforcement Learning Unresolved cited work
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 222c81a9-84bc-42e8-a0cf-3c4dc036e6f7 · outbound
Deep Transfer $Q$-Learning for Offline Non-Stationary Reinforcement Learning We now use the peeling argument to extend to uniform r
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 13b356e3-f83d-492e-8913-7039e61e011e · outbound
Deep Transfer $Q$-Learning for Offline Non-Stationary Reinforcement Learning For g1, · · ·, gN being an ϵ-covering set of G, we claim that g2 1 − eg2, · · ·, g2 N − eg2 is an 2bϵ-covering set of ¯G
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 393c7740-b859-4d7e-a972-8aecdb99fde4 · outbound
Deep Transfer $Q$-Learning for Offline Non-Stationary Reinforcement Learning In our dataset, the mortality rate is 24.21% for female and 22.71% for male
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 7e83532c-c609-4c66-9c6d-a8b9e2d5d5f1 · outbound
Deep Transfer $Q$-Learning for Offline Non-Stationary Reinforcement Learning Figure 5 in Chen, Li & Jordan (2022) presents mortality rates of different lengths
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 26bbd280-67f7-439d-b3c7-ea930d053232 · outbound
Deep Transfer $Q$-Learning for Offline Non-Stationary Reinforcement Learning B., Davidian, M
Reference 114
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation bba30a7c-7f58-49a8-8c84-50e975fb6d42 · outbound
Deep Transfer $Q$-Learning for Offline Non-Stationary Reinforcement Learning & Remlinger, C
Reference 343
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a019a3d5-0491-4d7b-becc-8ee33a89f447 · outbound
Deep Transfer $Q$-Learning for Offline Non-Stationary Reinforcement Learning & Song, R
Reference 640
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 523052cf-28b5-4f18-9179-f93475dfe160 · outbound
Deep Transfer $Q$-Learning for Offline Non-Stationary Reinforcement Learning Unresolved cited work
Reference 651
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 65c1489b-4bfd-445a-b7c5-d722c61da4b5 · outbound
Deep Transfer $Q$-Learning for Offline Non-Stationary Reinforcement Learning Pseudo-Labeling for Kernel Ridge Regression under Covariate Shift
Reference 901
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 70a0fd0f-6126-4ca7-8351-939c8091aa40 · outbound
Deep Transfer $Q$-Learning for Offline Non-Stationary Reinforcement Learning (2012), Transfer in reinforcement learning: A framework and a survey, in ‘Reinforcement Learning’, Springer, pp
Reference 1225
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 6f06d249-26eb-4209-8c26-c81945a89489 · outbound
Deep Transfer $Q$-Learning for Offline Non-Stationary Reinforcement Learning Deep Transfer Q-Learning for Offline Non-Stationary Reinforcement Learning
Reference 1549
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation d174b87b-2a88-4461-8cc9-ae65dc60a418 · outbound
Deep Transfer $Q$-Learning for Offline Non-Stationary Reinforcement Learning Unresolved cited work
Reference 2014
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation a80d550b-7804-4b87-a7ab-37bdeb6d35a5 · outbound
Deep Transfer $Q$-Learning for Offline Non-Stationary Reinforcement Learning On the Power of Multitask Representation Learning in Linear MDP
Reference 3364
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1c0b5eee-179a-49b9-aa6d-0e3978524b3f · inbound
Phase Transition in Nonparametric Minimax Rates for Covariate Shifts on Approximate Manifolds Deep Transfer $Q$-Learning for Offline Non-Stationary Reinforcement Learning
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3f7af179-139a-430a-b8a4-f82a314dddee · inbound
One-Step Bellman Alignment Enables Provably Efficient Transfer in Online RL Deep Transfer $Q$-Learning for Offline Non-Stationary Reinforcement Learning
Reference 2007
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d3d7ee4b-295b-4114-af5d-f0fd70ad5b99 · inbound
Kernelized Advantage Estimation: From Nonparametric Statistics to LLM Reasoning Deep Transfer $Q$-Learning for Offline Non-Stationary Reinforcement Learning
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 84e593e0-db75-4a6b-94c9-5b7a2ec81661 · inbound
Kernelized Advantage Estimation: From Nonparametric Statistics to LLM Reasoning Deep Transfer $Q$-Learning for Offline Non-Stationary Reinforcement Learning
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 5a3779c5-27b3-4cce-88ee-6f1bf801dfce · inbound
Dual-Channel Tensor Neural Networks: Finite-Sample Theory and Conformal Structure Selection Deep Transfer $Q$-Learning for Offline Non-Stationary Reinforcement Learning
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 15042302-ff5b-4033-9b07-7f2412602dbb · inbound
Learning to Hand Off: Provably Convergent Workflow Learning under Interface Constraints Deep Transfer $Q$-Learning for Offline Non-Stationary Reinforcement Learning
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.