Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-12T20:17:36.456856Z
Paper Citation Record · LEDGER
As of 15 August 2026, this Paper Citation Record lists 53 of 53 outbound references and 0 inbound Pith citation observations for arXiv:2411.09891.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-12T20:17:36.456856Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-15T06:32:42.880941+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
53 of 53 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 5613138e-a2b7-4f09-b075-fa23f26b74ca · outbound
Off-Dynamics Reinforcement Learning via Domain Adaptation and Reward Augmented Imitation Deep rein- forcement learning for dynamic treatment regimes on medical registry data
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 3ff031b4-4c2d-4c0a-9c54-4fa6922eb086 · outbound
Off-Dynamics Reinforcement Learning via Domain Adaptation and Reward Augmented Imitation Deep reinforcement learning for autonomous driving: A survey
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4baa2b41-6a6a-4dce-8b3a-f3a1fef53fbd · outbound
Off-Dynamics Reinforcement Learning via Domain Adaptation and Reward Augmented Imitation Off-Dynamics Reinforcement Learning: Training for Transfer with Domain Classifiers
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fa5f8d35-7bd6-4766-924c-660ddfb74b70 · outbound
Off-Dynamics Reinforcement Learning via Domain Adaptation and Reward Augmented Imitation Sim-to-real interactive recommendation via off-dynamics reinforcement learning
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation dd32523b-2d57-48a7-a4f0-b2c20d772568 · outbound
Off-Dynamics Reinforcement Learning via Domain Adaptation and Reward Augmented Imitation DARA: Dynamics-Aware Reward Augmentation in Offline Reinforcement Learning
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3a913f97-f674-4ba4-9dce-f911e3608368 · outbound
Off-Dynamics Reinforcement Learning via Domain Adaptation and Reward Augmented Imitation Unsupervised domain adaptation with dynamics-aware rewards in reinforcement learning
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 7499ddd5-1a1b-4f04-9a6f-ce7f3de53280 · outbound
Off-Dynamics Reinforcement Learning via Domain Adaptation and Reward Augmented Imitation Generative adversarial imitation learning
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fd30c84c-8a34-4fd0-8d82-19e9a3177f5a · outbound
Off-Dynamics Reinforcement Learning via Domain Adaptation and Reward Augmented Imitation Generative Adversarial Imitation from Observation
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 67525ab4-c36a-43f7-ba45-e7263eda8d37 · outbound
Off-Dynamics Reinforcement Learning via Domain Adaptation and Reward Augmented Imitation Offline imitation learning with a misspecified simulator
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 5fb528e4-6cfa-4290-a0b0-cef63400ccd0 · outbound
Off-Dynamics Reinforcement Learning via Domain Adaptation and Reward Augmented Imitation An imitation from observation approach to transfer learning with dynamics mismatch
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 2673544c-0942-4e83-b05d-3a617f6e9d54 · outbound
Off-Dynamics Reinforcement Learning via Domain Adaptation and Reward Augmented Imitation State-only Imitation with Transition Dynamics Mismatch
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 81241666-5130-4ff8-9e30-5624202aeeea · outbound
Off-Dynamics Reinforcement Learning via Domain Adaptation and Reward Augmented Imitation Doubly Robust Policy Evaluation and Learning
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5ff84bb0-e52e-42df-8718-052153b1fb00 · outbound
Off-Dynamics Reinforcement Learning via Domain Adaptation and Reward Augmented Imitation Doubly robust off-policy value evaluation for reinforcement learning
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 3fa4eb3f-6143-4c6c-b60b-5c3ca1ec149c · outbound
Off-Dynamics Reinforcement Learning via Domain Adaptation and Reward Augmented Imitation Doubly robust off-policy evaluation with shrinkage
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 93013354-fb64-4f17-a28a-0693ab6bfc36 · outbound
Off-Dynamics Reinforcement Learning via Domain Adaptation and Reward Augmented Imitation Doubly robust off-policy actor-critic: Convergence and optimality
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation a67cd4fa-e58b-40a3-a2ee-0b12438c2009 · outbound
Off-Dynamics Reinforcement Learning via Domain Adaptation and Reward Augmented Imitation Doubly robust distribu- tionally robust off-policy evaluation and learning
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 838e9822-b528-4ded-88ef-3632000b9f7e · outbound
Off-Dynamics Reinforcement Learning via Domain Adaptation and Reward Augmented Imitation Soft actor-critic: Off- policy maximum entropy deep reinforcement learning with a stochastic actor
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b521f84e-07f0-4a70-bd00-a9afb35559db · outbound
Off-Dynamics Reinforcement Learning via Domain Adaptation and Reward Augmented Imitation On the off-dynamics approach to reinforcement learning
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 983a1df7-5f2d-4b03-b13b-48f96b33386d · outbound
Off-Dynamics Reinforcement Learning via Domain Adaptation and Reward Augmented Imitation When to trust your model: Model-based policy optimization
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 2322f4d4-362a-4411-b663-48c7abd51685 · outbound
Off-Dynamics Reinforcement Learning via Domain Adaptation and Reward Augmented Imitation Mutual alignment transfer learning
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 859041a9-8267-4a5a-b3cc-e5099afa6451 · outbound
Off-Dynamics Reinforcement Learning via Domain Adaptation and Reward Augmented Imitation Domain Adaptation for Reinforcement Learning on the Atari
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation c9e7c01a-2133-406d-a598-39bc065876a7 · outbound
Off-Dynamics Reinforcement Learning via Domain Adaptation and Reward Augmented Imitation Domain adaptation in reinforcement learning via latent unified state representation
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 8df7979f-e2dd-4bd4-bc39-b81e5579f731 · outbound
Off-Dynamics Reinforcement Learning via Domain Adaptation and Reward Augmented Imitation Transfer learning in deep reinforce- ment learning: A survey
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 40b228c3-039a-44d7-b42e-dd51a0522d74 · outbound
Off-Dynamics Reinforcement Learning via Domain Adaptation and Reward Augmented Imitation Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1a0e4323-7e16-45d5-a78e-8ab3d648c68b · outbound
Off-Dynamics Reinforcement Learning via Domain Adaptation and Reward Augmented Imitation State regularized policy optimization on data with dynamics shift
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 3cc6b026-d4f9-4812-b364-efe5a0e14d5b · outbound
Off-Dynamics Reinforcement Learning via Domain Adaptation and Reward Augmented Imitation Generative adversarial nets
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 68384332-9525-40be-9946-9bd08d26d87c · outbound
Off-Dynamics Reinforcement Learning via Domain Adaptation and Reward Augmented Imitation Distributionally robust off-dynamics reinforcement learning: Prov- able efficiency with linear function approximation
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 41eaee32-c7f3-45e9-b0c1-aaa0782819e2 · outbound
Off-Dynamics Reinforcement Learning via Domain Adaptation and Reward Augmented Imitation Learning Robust Rewards with Adversarial Inverse Reinforcement Learning
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a118e89a-86dd-401d-bf96-580d44b86363 · outbound
Off-Dynamics Reinforcement Learning via Domain Adaptation and Reward Augmented Imitation Imitation learning via kernel mean embedding
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 8eb1bf54-3c9e-42ae-a0e2-dbb200a85835 · outbound
Off-Dynamics Reinforcement Learning via Domain Adaptation and Reward Augmented Imitation Variational Discriminator Bottleneck: Improving Imitation Learning, Inverse RL, and GANs by Constraining Information Flow
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6f4a63f3-1e08-4d69-9b0b-6e5b06b5f65b · outbound
Off-Dynamics Reinforcement Learning via Domain Adaptation and Reward Augmented Imitation Task transfer by preference-based cost learning
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation c26186d0-e667-49f6-8204-4dc139411b68 · outbound
Off-Dynamics Reinforcement Learning via Domain Adaptation and Reward Augmented Imitation Imitation Learning from Video by Leveraging Proprioception
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 16e2bad1-a861-4f73-a727-725b0cec41eb · outbound
Off-Dynamics Reinforcement Learning via Domain Adaptation and Reward Augmented Imitation Imitation from observation: Learning to imitate behaviors from raw video via context translation
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation e3e99ab7-cf1f-401e-ade8-3e44214fb1a7 · outbound
Off-Dynamics Reinforcement Learning via Domain Adaptation and Reward Augmented Imitation Behavioral Cloning from Observation
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cd31525b-a218-4e47-b5c5-196c1eb8eb77 · outbound
Off-Dynamics Reinforcement Learning via Domain Adaptation and Reward Augmented Imitation Recent Advances in Imitation Learning from Observation
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b555e714-d3e7-4ea0-92ec-8fa8f3e60266 · outbound
Off-Dynamics Reinforcement Learning via Domain Adaptation and Reward Augmented Imitation Domain adaptive imitation learning
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 99553806-6256-4bc9-89a8-cd82d7a4e646 · outbound
Off-Dynamics Reinforcement Learning via Domain Adaptation and Reward Augmented Imitation Generalization and equilibrium in generative adversarial nets (gans)
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 961c3012-d6c7-4f41-ab91-36a32d94b33b · outbound
Off-Dynamics Reinforcement Learning via Domain Adaptation and Reward Augmented Imitation vf+MOWCXlXkD7CB/Zj9tqm3hyT0=
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation d54b067f-39d8-4ab1-bdba-831510ef0bcf · outbound
Off-Dynamics Reinforcement Learning via Domain Adaptation and Reward Augmented Imitation And in the introduc- tion section, we have a contribution list
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation a1a4d7ba-c253-4094-ab40-5eb181ccc063 · outbound
Off-Dynamics Reinforcement Learning via Domain Adaptation and Reward Augmented Imitation Limitations
Reference 40
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 0155ba99-1d8a-4638-aa00-a973529a342e · outbound
Off-Dynamics Reinforcement Learning via Domain Adaptation and Reward Augmented Imitation We present our theoretical result in Section 4 and the proof is in Appendix B
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation b0cdfb74-91ad-4fcd-9254-ad5b9abc7f1f · outbound
Off-Dynamics Reinforcement Learning via Domain Adaptation and Reward Augmented Imitation Guidelines: • The answer NA means that the paper does not include experiments
Reference 42
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 33486d2d-81cc-46f8-985c-89e0480a87b9 · outbound
Off-Dynamics Reinforcement Learning via Domain Adaptation and Reward Augmented Imitation Guidelines: • The answer NA means that paper does not include experiments requiring code
Reference 43
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 7601504f-19b7-48d6-abc1-8400b07369ed · outbound
Off-Dynamics Reinforcement Learning via Domain Adaptation and Reward Augmented Imitation We also describe the hyperparameter tuning in the Appendix D.4
Reference 44
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation e3f0739b-d912-422c-815e-64f036c38ccd · outbound
Off-Dynamics Reinforcement Learning via Domain Adaptation and Reward Augmented Imitation Guidelines: • The answer NA means that the paper does not include experiments
Reference 45
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 6c8fae34-bc3a-4d87-8a50-d06a4c9b1a9b · outbound
Off-Dynamics Reinforcement Learning via Domain Adaptation and Reward Augmented Imitation Guidelines: • The answer NA means that the paper does not include experiments
Reference 46
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 03c7ca5f-168a-4789-8b6d-8373cf47c48e · outbound
Off-Dynamics Reinforcement Learning via Domain Adaptation and Reward Augmented Imitation Guidelines: • The answer NA means that the authors have not reviewed the NeurIPS Code of Ethics
Reference 47
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 9398912b-84f5-473b-8888-9b4764310394 · outbound
Off-Dynamics Reinforcement Learning via Domain Adaptation and Reward Augmented Imitation Guidelines: • The answer NA means that there is no societal impact of the work performed
Reference 48
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 9b412b4b-d424-486a-b262-bf11491823fe · outbound
Off-Dynamics Reinforcement Learning via Domain Adaptation and Reward Augmented Imitation Guidelines: • The answer NA means that the paper poses no such risks
Reference 49
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation eccf6cce-6e09-4549-8514-24227f4aff22 · outbound
Off-Dynamics Reinforcement Learning via Domain Adaptation and Reward Augmented Imitation Guidelines: • The answer NA means that the paper does not use existing assets
Reference 50
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation aadab1a8-e210-4843-bc9d-b1dcb1ae1500 · outbound
Off-Dynamics Reinforcement Learning via Domain Adaptation and Reward Augmented Imitation Also, details about the implementation are included in the paper
Reference 51
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 44774b8e-2c5d-4bd7-b370-8cea819e8347 · outbound
Off-Dynamics Reinforcement Learning via Domain Adaptation and Reward Augmented Imitation Guidelines: • The answer NA means that the paper does not involve crowdsourcing nor research with human subjects
Reference 52
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation c3afa81f-16bb-4e33-b2b4-90db56f1ed3d · outbound
Off-Dynamics Reinforcement Learning via Domain Adaptation and Reward Augmented Imitation Guidelines: • The answer NA means that the paper does not involve crowdsourcing nor research with human subjects
Reference 53
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
No inbound Pith citation observations are available.