Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T14:26:53.058645Z
Paper Citation Record · LEDGER
As of 8 August 2026, this Paper Citation Record lists 66 of 66 outbound references and 2 inbound Pith citation observations for arXiv:2505.19058.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T14:26:53.058645Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-07-10T09:56:20.351339Z
A source-named dated measurement, never combined with another source.
Source: pith, observed 2026-07-10T09:57:00.793817Z
66 of 66 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation d56f98fb-9fae-4388-a4fb-338b43c8f91d · outbound
Distributionally Robust Deep Q-Learning Investigating the parameters of the beta distribution
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 9ac03d49-301b-462e-aa0e-f29bd48cd45e · outbound
Distributionally Robust Deep Q-Learning Infinite dimensional analysis: a hitchhiker’s guide
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation b7f522da-e813-468e-a6b7-fbdbc4005c9e · outbound
Distributionally Robust Deep Q-Learning Computational aspects of robust optimized certainty equiv- alents and option pricing
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation c17de44c-bf89-4428-b2f5-f41e89f22035 · outbound
Distributionally Robust Deep Q-Learning On the theory of dynamic programming
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation ae0c3ccc-57a8-4ac0-b321-2fa2e0c1c04d · outbound
Distributionally Robust Deep Q-Learning Dota 2 with Large Scale Deep Reinforcement Learning
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fc8655b3-0994-46f6-b078-75faa5cf58a5 · outbound
Distributionally Robust Deep Q-Learning Convex optimization
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 335e8b6a-d2cc-430c-a794-e6c70c8e712c · outbound
Distributionally Robust Deep Q-Learning Distributionally robust Markov decision processes and their connection to risk measures
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 1461b9f8-e872-4b1c-8a58-f98767d84ecf · outbound
Distributionally Robust Deep Q-Learning Unresolved cited work
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation c8bb8b23-9445-48d0-bd0c-0f859f07c0a6 · outbound
Distributionally Robust Deep Q-Learning Sinkhorn distances: Lightspeed computation of optimal transport, 2013
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 836e6293-d9a6-4848-a7b4-c4090a822da9 · outbound
Distributionally Robust Deep Q-Learning Robust Q-learning for finite ambiguity sets
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 757a9a1a-eb1b-4fc5-b04e-5637ca04f287 · outbound
Distributionally Robust Deep Q-Learning Twice Regularized Markov Decision Processes: The Equivalence between Robustness and Regularization
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 016885b5-6336-4a61-83bc-6130eab112a1 · outbound
Distributionally Robust Deep Q-Learning Maximum Entropy RL (Provably) Solves Some Robust RL Problems
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6ebf42a6-9e34-4417-be0d-d482898157cd · outbound
Distributionally Robust Deep Q-Learning A theoretical analysis of deep Q-learning
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 383dc3d1-25d8-4bb0-b7dc-22b3084143b2 · outbound
Distributionally Robust Deep Q-Learning In- terpolating between optimal transport and mmd using Sinkhorn divergences
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 742e4c8e-f8e8-491c-8a58-6d57f19be4c7 · outbound
Distributionally Robust Deep Q-Learning Sample complexity of Sinkhorn divergences
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 61f552e5-ee19-46ee-931f-1072e7e14675 · outbound
Distributionally Robust Deep Q-Learning Stability of entropic optimal transport and Schr¨ odinger bridges
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 80cc6a39-69e5-4015-b210-74594bae82a9 · outbound
Distributionally Robust Deep Q-Learning Robust Markov decision processes: Beyond rectangularity.Mathematics of Operations Research, 48(1):203–226, 2023
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation f6bfc4ed-785b-4e05-9a94-c43e1aedc790 · outbound
Distributionally Robust Deep Q-Learning Double Q-learning
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 14cce011-4b45-4032-9112-e2dbd88ecc1b · outbound
Distributionally Robust Deep Q-Learning Universal approximation of an unknown mapping and its derivatives using multilayer feedforward networks
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b10a21ee-df77-4096-b7eb-947881c8ce83 · outbound
Distributionally Robust Deep Q-Learning Learning to utilize shaping rewards: A new approach of reward shaping
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c468e487-5adc-4025-a1aa-9d4ca6407773 · outbound
Distributionally Robust Deep Q-Learning Robust dynamic programming
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 54768045-e007-48f4-b640-1e1cf17b974a · outbound
Distributionally Robust Deep Q-Learning Probability essentials
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 44bda2c0-d2e0-465a-95df-957062227227 · outbound
Distributionally Robust Deep Q-Learning Universal approximation with deep narrow networks
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5daadba4-6eb4-4020-a685-1df0d1095287 · outbound
Distributionally Robust Deep Q-Learning Probability theory: a comprehensive course
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 2196972d-d2b4-4108-97d5-1b9403556f1d · outbound
Distributionally Robust Deep Q-Learning An Efficient Solution to s-Rectangular Robust Markov Decision Processes
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 7514de8d-3eea-48f6-971f-ef12ca6fd1ec · outbound
Distributionally Robust Deep Q-Learning Playing fps games with deep reinforcement learning
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation eb124f69-56be-46e8-bf84-c9cdfc65cddb · outbound
Distributionally Robust Deep Q-Learning Policy gradient algorithms for robust MDPs with non-rectangular uncertainty sets
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cc11521d-e1d2-4be7-b7f0-b84df17fb853 · outbound
Distributionally Robust Deep Q-Learning On the efficiency of entropic regularized algorithms for optimal transport
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation c7092ac4-6b7b-4725-93a5-afb950f7fa7c · outbound
Distributionally Robust Deep Q-Learning Distri- butionally robust Q-learning
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 2e2e4b44-ad3f-4779-b02b-8752d1864697 · outbound
Distributionally Robust Deep Q-Learning Generative model for financial time series trained with MMD using a signature kernel
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e7ebd7b1-035f-4e75-9410-b57014f7b327 · outbound
Distributionally Robust Deep Q-Learning Robust MDPs with k-rectangular uncertainty.Mathematics of Operations Research, 41(4):1484–1509, 2016
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 8e1c0b4b-89f1-4d19-ae29-31fa38fe1ceb · outbound
Distributionally Robust Deep Q-Learning Rusu, Joel Veness, Marc G
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation d33f3249-3bb9-4739-af78-1e972bea358f · outbound
Distributionally Robust Deep Q-Learning Robust SGLD algorithm for solving non-convex distributionally robust optimisation problems
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation de0a5a18-4943-49b3-87de-d696a478000b · outbound
Distributionally Robust Deep Q-Learning Universal approximation results for neural networks with non-polynomial activation function over non-compact domains
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c1a5b965-ff5d-465a-bfc0-3bcb47d4fa5e · outbound
Distributionally Robust Deep Q-Learning Neural networks can detect model-free static arbitrage strategies
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 8042d1d6-830d-4a55-9209-fbb9db4e4e27 · outbound
Distributionally Robust Deep Q-Learning Robust Q-learning algorithm for markov decision processes under Wasserstein uncertainty
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation a298fc5b-8367-49f9-a359-13bf6d74ecac · outbound
Distributionally Robust Deep Q-Learning Non-concave stochastic optimal control in finite discrete time under model uncertainty
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation d9cd9e9c-8e3a-42e0-a377-cd107430d60b · outbound
Distributionally Robust Deep Q-Learning Markov decision processes under model uncertainty.Mathematical Finance, 33(3):618–665, 2023
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 2afa2ddc-496e-4315-bb2f-1cf669188761 · outbound
Distributionally Robust Deep Q-Learning Robust control of Markov decision processes with uncertain transition matrices
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation f2560ea4-7670-4b2b-a2bf-adbff400a0d5 · outbound
Distributionally Robust Deep Q-Learning Introduction to entropic optimal transport
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b2049884-5728-4488-b3a3-ba753efab79c · outbound
Distributionally Robust Deep Q-Learning Robust reinforcement learning using offline data
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation eea69139-5433-4db7-a707-11bffd83a295 · outbound
Distributionally Robust Deep Q-Learning Approximation theory of the MLP model in neural networks
Reference 42
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation ddba481d-1872-4c3d-8d0d-08a0d80b2ec0 · outbound
Distributionally Robust Deep Q-Learning Distributionally Robust Optimization: A Review
Reference 43
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8cbb9b91-566c-43a2-b474-e6d9ec509828 · outbound
Distributionally Robust Deep Q-Learning Distributionally robust model-based reinforcement learning with large state spaces
Reference 44
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation c1040ac8-a747-415d-883d-d2afbf624f0f · outbound
Distributionally Robust Deep Q-Learning Principles of mathematical analysis , volume 3
Reference 45
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation e4022c2a-03d7-4696-aefe-c7f312d7cf30 · outbound
Distributionally Robust Deep Q-Learning Structural estimation of Markov decision processes
Reference 46
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation de461997-29b7-4fbc-b4f7-2c936034e59a · outbound
Distributionally Robust Deep Q-Learning Universal approximation using feedforward neural networks: A survey of some existing methods, and some new results
Reference 47
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation e5d9e8fa-928a-4bf0-bce8-48391be90837 · outbound
Distributionally Robust Deep Q-Learning A relationship between arbitrary positive matrices and doubly stochastic matrices
Reference 48
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 66540415-4f95-4bf3-ab46-6545ddfb849a · outbound
Distributionally Robust Deep Q-Learning Distributionally Robust Reinforcement Learning
Reference 49
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 94800049-d17f-4781-8b6f-e4ea6e729172 · outbound
Distributionally Robust Deep Q-Learning Reinforcement learning: An introduction
Reference 50
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation daddacaa-5731-4c90-af39-38eaa04cb3cb · outbound
Distributionally Robust Deep Q-Learning Sinkhorn Divergences for Unbalanced Optimal Transport
Reference 51
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 90a94789-bcdc-4323-8563-707e0907b41c · outbound
Distributionally Robust Deep Q-Learning Deep reinforcement learning: From Q-learning to deep Q-learning
Reference 52
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 9eb8acb0-7ace-4c7a-a575-38fa53266750 · outbound
Distributionally Robust Deep Q-Learning Deep reinforcement learning with double Q-learning
Reference 53
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation c2722181-b1d1-4d24-a021-bab0144a683b · outbound
Distributionally Robust Deep Q-Learning Springer, 2009
Reference 54
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 8a7301d0-28a0-4060-9ebc-8fe544c6c5b7 · outbound
Distributionally Robust Deep Q-Learning Sinkhorn Distributionally Robust Optimization
Reference 55
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d1c9673b-498d-493f-aa91-e2d8a0a25d65 · outbound
Distributionally Robust Deep Q-Learning Policy gradient in robust MDPs with global convergence guarantee, 2023
Reference 56
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 5a9ea900-7541-46ae-8ad3-2da91d988d08 · outbound
Distributionally Robust Deep Q-Learning A finite sample complexity bound for distribu- tionally robust Q-learning
Reference 57
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 1fdd2e0b-7b1d-4ce1-9e14-8bbca4e226fd · outbound
Distributionally Robust Deep Q-Learning Sample complexity of variance-reduced distribu- tionally robust Q-learning
Reference 58
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation e46641e0-b583-4a1a-a953-2f322065ca27 · outbound
Distributionally Robust Deep Q-Learning Online robust reinforcement learning with model uncertainty
Reference 59
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a4a41a19-4422-4fd5-ae3a-fdaa96631e48 · outbound
Distributionally Robust Deep Q-Learning Policy gradient method for robust reinforcement learning, 2022
Reference 60
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 0848a997-d20d-44f5-984c-55c6e2a6e9e0 · outbound
Distributionally Robust Deep Q-Learning Unresolved cited work
Reference 61
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 33fe3bf7-a23e-473b-bd21-36faa5913b7a · outbound
Distributionally Robust Deep Q-Learning Robust Markov decision processes
Reference 62
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 1b3e7aef-70fb-45c3-88bd-798f2076fa29 · outbound
Distributionally Robust Deep Q-Learning Distributionally robust Markov decision processes
Reference 63
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation fc240353-9188-43fd-8ea5-ac5677b2b03d · outbound
Distributionally Robust Deep Q-Learning A convex optimization approach to distributionally robust Markov decision processes with Wasser- stein distance
Reference 64
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 8158a332-f577-4700-bfb1-88cf205a7ebc · outbound
Distributionally Robust Deep Q-Learning Wasserstein distributionally robust stochastic control: A data-driven approach
Reference 65
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 1e5f6c2c-bf72-44e7-87f3-7a550d9b32ce · outbound
Distributionally Robust Deep Q-Learning On linear optimization over Wasserstein balls
Reference 66
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 1aaa4507-cffe-4cfc-8c0a-8764e250a890 · inbound
Robust $Q$-learning for mean-field control under Wasserstein uncertainty in common noise Distributionally Robust Deep Q-Learning
Reference 61
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 8e06f30b-2c3b-48d9-8ffd-d411166a95ae · inbound
Robustness in Sequential Decision Making under Evolving Uncertainty: Evidence from High-Frequency Market Making Distributionally Robust Deep Q-Learning
Reference 105
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.