Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-06T19:19:29.466784Z
Paper Citation Record · LEDGER
As of 7 August 2026, this Paper Citation Record lists 94 of 94 outbound references and 2 inbound Pith citation observations for arXiv:2507.06111.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-06T19:19:29.466784Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-06-28T22:51:20.005562Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-06-28T22:52:44.957699Z
94 of 94 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 1526371c-a675-4c97-814b-746330775ae9 · outbound
Safe Domain Randomization via Uncertainty-Aware Out-of-Distribution Detection and Policy Adaptation MIT press, 2018
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4bf1724f-2714-4f22-86e8-f3e3a134e07f · outbound
Safe Domain Randomization via Uncertainty-Aware Out-of-Distribution Detection and Policy Adaptation Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ba244a59-48ea-448b-a35c-4ea07b740dff · outbound
Safe Domain Randomization via Uncertainty-Aware Out-of-Distribution Detection and Policy Adaptation Reinforcement learning in robotics: A survey
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 56859d3c-2cd0-4236-95af-df70614d7519 · outbound
Safe Domain Randomization via Uncertainty-Aware Out-of-Distribution Detection and Policy Adaptation Toward self-driving processes: A deep reinforcement learning approach to control.AIChE journal, 65(10):e16689, 2019
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5073135a-bf34-45ec-9085-8493b5d7d39f · outbound
Safe Domain Randomization via Uncertainty-Aware Out-of-Distribution Detection and Policy Adaptation Sim-to-real transfer in deep reinforcement learning for robotics: a survey
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 15a92a1b-465b-48e0-bb1d-e67a041c145d · outbound
Safe Domain Randomization via Uncertainty-Aware Out-of-Distribution Detection and Policy Adaptation Out-of-Distribution Dynamics Detection: RL-Relevant Benchmarks and Results
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 21553616-b72f-42f5-9ff6-d9b1a9d7554f · outbound
Safe Domain Randomization via Uncertainty-Aware Out-of-Distribution Detection and Policy Adaptation Off-dynamics reinforcement learning: Training for transfer with domain classifiers
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation dbbeceac-cdb8-4aa8-9620-a67d693c4e8a · outbound
Safe Domain Randomization via Uncertainty-Aware Out-of-Distribution Detection and Policy Adaptation Robust dynamic programming.Mathematics of Operations Research, 30(2):257–280, 2005
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 028831bd-968d-4d86-babd-2a5893be2385 · outbound
Safe Domain Randomization via Uncertainty-Aware Out-of-Distribution Detection and Policy Adaptation Offline-to-online reinforcement learning via balanced replay and pessimistic q-ensemble
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2da0939b-9d49-4471-b2dc-9fccee67e3ac · outbound
Safe Domain Randomization via Uncertainty-Aware Out-of-Distribution Detection and Policy Adaptation Adaptive policy learning for offline-to-online reinforcement learning
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 22ed0b8b-7e03-4621-a290-1cf07f5008e1 · outbound
Safe Domain Randomization via Uncertainty-Aware Out-of-Distribution Detection and Policy Adaptation Domain randomization for transferring deep neural networks from simulation to the real world
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0b98c3b2-914a-4c4f-a18a-98557b79b8e0 · outbound
Safe Domain Randomization via Uncertainty-Aware Out-of-Distribution Detection and Policy Adaptation Pal, and Liam Paull
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 52067f97-ed3b-4a8f-9b0a-32e2eed5252f · outbound
Safe Domain Randomization via Uncertainty-Aware Out-of-Distribution Detection and Policy Adaptation Mujoco: A physics engine for model-based control
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2db63577-4955-4e25-b994-5b2c2564a860 · outbound
Safe Domain Randomization via Uncertainty-Aware Out-of-Distribution Detection and Policy Adaptation Towards a generic solution for inspection of industrial sites
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bc1b65b0-5a98-4c53-8587-fff94163a914 · outbound
Safe Domain Randomization via Uncertainty-Aware Out-of-Distribution Detection and Policy Adaptation Offline reinforcement learning with implicit q-learning
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5bb02de6-6422-450c-8a4b-b8bf88b0c20b · outbound
Safe Domain Randomization via Uncertainty-Aware Out-of-Distribution Detection and Policy Adaptation Way Off-Policy Batch Deep Reinforcement Learning of Implicit Human Preferences in Dialog
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ee08bffa-b9a0-4338-8422-6d1c195222ae · outbound
Safe Domain Randomization via Uncertainty-Aware Out-of-Distribution Detection and Policy Adaptation A minimalist approach to offline reinforcement learning
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bee2bd30-27bc-4279-ae5d-b52113dc3e41 · outbound
Safe Domain Randomization via Uncertainty-Aware Out-of-Distribution Detection and Policy Adaptation Conservative q-learning for offline reinforcement learning
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 20eb6e39-95f7-43eb-a934-3d10f6fbe10b · outbound
Safe Domain Randomization via Uncertainty-Aware Out-of-Distribution Detection and Policy Adaptation Uncertainty-based of- fline reinforcement learning with diversified q-ensemble
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation c53f3ea2-6348-499e-a952-2eb398ca4f21 · outbound
Safe Domain Randomization via Uncertainty-Aware Out-of-Distribution Detection and Policy Adaptation Iteratively refined behavior regularization for offline reinforcement learning
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation ee8242d5-9d9d-4fff-a90a-9e4cc5f1d450 · outbound
Safe Domain Randomization via Uncertainty-Aware Out-of-Distribution Detection and Policy Adaptation Offline reinforcement learning with OOD state correction and OOD action suppression
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 0a9c719c-02dd-45dd-879a-57eafe701b56 · outbound
Safe Domain Randomization via Uncertainty-Aware Out-of-Distribution Detection and Policy Adaptation Model-Bellman inconsistency for model-based offline reinforcement learning
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8048c12b-3fb2-4016-9c99-9a59bc06f200 · outbound
Safe Domain Randomization via Uncertainty-Aware Out-of-Distribution Detection and Policy Adaptation Pessimistic bootstrapping for uncertainty-driven offline reinforcement learning
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 6d05255c-a63e-4185-8418-3899e31904ee · outbound
Safe Domain Randomization via Uncertainty-Aware Out-of-Distribution Detection and Policy Adaptation Rorl: Ro- bust offline reinforcement learning via conservative smoothing.Advances in neural information processing systems, 35:23851–23866, 2022
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 0021822e-187f-4dc2-aa2a-15414400336e · outbound
Safe Domain Randomization via Uncertainty-Aware Out-of-Distribution Detection and Policy Adaptation When to trust your simulator: Dynamics-aware hybrid offline-and-online reinforcement learning
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 1648d827-69cd-4abd-866e-707f3e859cda · outbound
Safe Domain Randomization via Uncertainty-Aware Out-of-Distribution Detection and Policy Adaptation Cross-domain policy adaptation via value-guided data filtering
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 5cad286d-7c02-4cb0-a929-756ec34e37be · outbound
Safe Domain Randomization via Uncertainty-Aware Out-of-Distribution Detection and Policy Adaptation Cross-domain policy adaptation by capturing representation mismatch
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation b8338a00-9f6c-41d4-b8ef-7e2fa00fc332 · outbound
Safe Domain Randomization via Uncertainty-Aware Out-of-Distribution Detection and Policy Adaptation Unsolved Problems in ML Safety
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 20628592-55e8-4951-a6f7-9e72589de342 · outbound
Safe Domain Randomization via Uncertainty-Aware Out-of-Distribution Detection and Policy Adaptation Learning dexterous in-hand manipulation.The International Journal of Robotics Research, 39(1):3–20, 2020
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ff54dac0-ddf8-4970-a069-6a11f9a2fc24 · outbound
Safe Domain Randomization via Uncertainty-Aware Out-of-Distribution Detection and Policy Adaptation Network randomization: A simple technique for generalization in deep reinforcement learning
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 135a6683-631b-424c-98a9-34fcc6649337 · outbound
Safe Domain Randomization via Uncertainty-Aware Out-of-Distribution Detection and Policy Adaptation Learning domain randomization distributions for training robust locomotion policies
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation a7568781-4dda-4914-a03b-96015ec8da77 · outbound
Safe Domain Randomization via Uncertainty-Aware Out-of-Distribution Detection and Policy Adaptation A Markovian Decision Process.Indiana University Mathematics Journal, 6(4):679–684, 1957
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 6ea5ee09-e2c6-453c-abae-4c239af013f0 · outbound
Safe Domain Randomization via Uncertainty-Aware Out-of-Distribution Detection and Policy Adaptation Handling black swan events in deep learning with diversely extrapolated neural networks
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 3d72acdb-e36d-42a2-ab81-a9edbd72c976 · outbound
Safe Domain Randomization via Uncertainty-Aware Out-of-Distribution Detection and Policy Adaptation Cal-QL: Calibrated offline RL pre-training for efficient online fine-tuning
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 89de8493-1f4e-4a75-a416-b3beae587f0d · outbound
Safe Domain Randomization via Uncertainty-Aware Out-of-Distribution Detection and Policy Adaptation AWAC: Accelerating Online Reinforcement Learning with Offline Datasets
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 465d13c9-be27-4ae0-88e8-a76087b78458 · outbound
Safe Domain Randomization via Uncertainty-Aware Out-of-Distribution Detection and Policy Adaptation D4RL: Datasets for Deep Data-Driven Reinforcement Learning
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9b219c75-4ffa-4510-a637-a87ed8da06a3 · outbound
Safe Domain Randomization via Uncertainty-Aware Out-of-Distribution Detection and Policy Adaptation Terry, Ariel Kwiatkowski, John U
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation e0d11664-5751-476a-9d5d-27ba11d0f71b · outbound
Safe Domain Randomization via Uncertainty-Aware Out-of-Distribution Detection and Policy Adaptation Soft actor-critic: Off- policy maximum entropy deep reinforcement learning with a stochastic actor
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation d9b9f682-7946-44a3-9028-18bc64fda4c4 · outbound
Safe Domain Randomization via Uncertainty-Aware Out-of-Distribution Detection and Policy Adaptation Addressing function approximation error in actor-critic methods
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b1d777a0-3c31-41ae-9faa-2890d9c51f4f · outbound
Safe Domain Randomization via Uncertainty-Aware Out-of-Distribution Detection and Policy Adaptation Corl: Research-oriented deep offline reinforcement learning library
Reference 40
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 752d14ce-4484-46ae-bca1-90b40391f100 · outbound
Safe Domain Randomization via Uncertainty-Aware Out-of-Distribution Detection and Policy Adaptation Efficient online reinforcement learning with offline data
Reference 41
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7a3f1e80-44d1-4aec-bad1-b6334b7b664a · outbound
Safe Domain Randomization via Uncertainty-Aware Out-of-Distribution Detection and Policy Adaptation Odrl: A benchmark for off-dynamics reinforcement learning
Reference 42
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation cb7d1f36-c7fa-42cf-87d0-45e79e5fbf57 · outbound
Safe Domain Randomization via Uncertainty-Aware Out-of-Distribution Detection and Policy Adaptation Darl: distance-aware uncertainty estimation for offline reinforcement learning
Reference 43
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation f625135f-81d6-4579-9c72-d976d15e5d46 · outbound
Safe Domain Randomization via Uncertainty-Aware Out-of-Distribution Detection and Policy Adaptation Open RL Benchmark: Comprehensive Tracked Experiments for Reinforcement Learning
Reference 44
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c4ce109d-cc73-4469-8485-8f5f7a86c5a1 · outbound
Safe Domain Randomization via Uncertainty-Aware Out-of-Distribution Detection and Policy Adaptation Dario Bellicoso, Vassilios Tsounis, Jemin Hwangbo, Karen Bodie, Peter Fankhauser, Michael Bloesch, Remo Diethelm, Samuel Bachmann, Amir Melzer, and Mark Hoepflinger
Reference 45
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 50084dd9-436a-4881-bbcc-a0a01590d0ff · outbound
Safe Domain Randomization via Uncertainty-Aware Out-of-Distribution Detection and Policy Adaptation Learning to walk in minutes using massively parallel deep reinforcement learning
Reference 46
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation d4297451-6beb-4138-ba31-b904e8cbec12 · outbound
Safe Domain Randomization via Uncertainty-Aware Out-of-Distribution Detection and Policy Adaptation Learning agile and dynamic motor skills for legged robots.Science Robotics, 4(26):eaau5872, 2019
Reference 47
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9aaf98a9-3769-4d8f-abf7-4fa1ce272a68 · outbound
Safe Domain Randomization via Uncertainty-Aware Out-of-Distribution Detection and Policy Adaptation REvolveR: Continuous evolutionary models for robot-to-robot policy transfer
Reference 48
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation a8f851a0-5087-40e1-879a-1118306b0361 · outbound
Safe Domain Randomization via Uncertainty-Aware Out-of-Distribution Detection and Policy Adaptation PROTO: Iterative Policy Regularized Offline-to-Online Reinforcement Learning
Reference 49
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 819b9590-101f-4542-b17c-1bcf0a29a9bd · outbound
Safe Domain Randomization via Uncertainty-Aware Out-of-Distribution Detection and Policy Adaptation Off-policy deep reinforcement learning without exploration
Reference 50
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation f0f42879-8418-4202-aa5e-d81b0f110029 · outbound
Safe Domain Randomization via Uncertainty-Aware Out-of-Distribution Detection and Policy Adaptation An optimistic perspective on offline reinforcement learning
Reference 51
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 29ac67b7-ce28-40b4-8b7b-21d4cd3a7124 · outbound
Safe Domain Randomization via Uncertainty-Aware Out-of-Distribution Detection and Policy Adaptation Tree-based batch mode reinforcement learning.Journal of Machine Learning Research, 6, 2005
Reference 52
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 8dfcd02c-6f0d-40ba-add7-c9cfdb0124d1 · outbound
Safe Domain Randomization via Uncertainty-Aware Out-of-Distribution Detection and Policy Adaptation Stabilizing off-policy q-learning via bootstrapping error reduction
Reference 53
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 55874838-495d-4e38-bb63-cfb272151dc6 · outbound
Safe Domain Randomization via Uncertainty-Aware Out-of-Distribution Detection and Policy Adaptation Batch reinforcement learning
Reference 54
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation e0640f26-8c90-42d9-bbf1-2ffbe8ff7bd0 · outbound
Safe Domain Randomization via Uncertainty-Aware Out-of-Distribution Detection and Policy Adaptation Critic regularized regression.Advances in Neural Information Processing Systems, 33:7768–7778, 2020
Reference 55
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation f78f2176-cf11-4722-8aea-42ac0e26f1ad · outbound
Safe Domain Randomization via Uncertainty-Aware Out-of-Distribution Detection and Policy Adaptation Revisiting the minimalist approach to offline reinforcement learning.Advances in Neural Information Processing Systems, 36, 2024
Reference 56
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation f2b04cf0-7311-4c3e-92b0-5e76903e2726 · outbound
Safe Domain Randomization via Uncertainty-Aware Out-of-Distribution Detection and Policy Adaptation Behavior Regularized Offline Reinforcement Learning
Reference 57
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1b2f9574-4016-42a6-bb3e-120b67be6ac6 · outbound
Safe Domain Randomization via Uncertainty-Aware Out-of-Distribution Detection and Policy Adaptation Keep doing what worked: Behavior modelling priors for offline reinforcement learning
Reference 58
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation bd2927dd-016b-4e78-aed0-376aab5293fa · outbound
Safe Domain Randomization via Uncertainty-Aware Out-of-Distribution Detection and Policy Adaptation Offline reinforcement learning with fisher divergence critic regularization
Reference 59
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation ccdaffb5-16a9-467c-9476-671ed0365b51 · outbound
Safe Domain Randomization via Uncertainty-Aware Out-of-Distribution Detection and Policy Adaptation Uncertainty weighted actor-critic for offline reinforcement learning
Reference 60
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 89e1fb23-82cc-48ab-8963-ad7587a9de85 · outbound
Safe Domain Randomization via Uncertainty-Aware Out-of-Distribution Detection and Policy Adaptation Uni-o4: Unifying online and offline deep reinforcement learning with multi-step on-policy optimization
Reference 61
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 963958b2-a71a-488e-8987-d9889d18be8f · outbound
Safe Domain Randomization via Uncertainty-Aware Out-of-Distribution Detection and Policy Adaptation Enoto: Improving offline-to-online reinforcement learning with q-ensembles
Reference 62
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation ab3132ae-9778-49a4-84b9-319bea955ec4 · outbound
Safe Domain Randomization via Uncertainty-Aware Out-of-Distribution Detection and Policy Adaptation Online decision transformer
Reference 63
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation fb53cca9-2672-4a72-bd95-01c21239d0cb · outbound
Safe Domain Randomization via Uncertainty-Aware Out-of-Distribution Detection and Policy Adaptation Actor-critic alignment for offline-to-online reinforcement learning
Reference 64
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 1df1ff13-ab37-47fc-b576-1c1c66aaeb7e · outbound
Safe Domain Randomization via Uncertainty-Aware Out-of-Distribution Detection and Policy Adaptation Train once, get a family: State-adaptive balances for offline-to-online reinforcement learning.Advances in Neural Information Processing Systems, 36, 2024
Reference 65
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 6e2dcd07-f1b7-45e6-b59d-b4f3974fbdbc · outbound
Safe Domain Randomization via Uncertainty-Aware Out-of-Distribution Detection and Policy Adaptation Policy expansion for bridging offline-to-online reinforcement learning
Reference 66
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 9faaf102-b410-4ae3-bf10-5977dcb8ff12 · outbound
Safe Domain Randomization via Uncertainty-Aware Out-of-Distribution Detection and Policy Adaptation Albrecht, and Amos Storkey
Reference 67
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation aeff3033-7800-4f74-a7f8-cfbab583a70e · outbound
Safe Domain Randomization via Uncertainty-Aware Out-of-Distribution Detection and Policy Adaptation A comprehensive survey on safe reinforcement learning
Reference 68
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 866b0f2c-0144-4aa1-857d-504538951ec7 · outbound
Safe Domain Randomization via Uncertainty-Aware Out-of-Distribution Detection and Policy Adaptation Consideration of risk in reinforcement learning
Reference 69
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation ad5a7b7e-cd11-4111-a7a1-b6e87ac55b02 · outbound
Safe Domain Randomization via Uncertainty-Aware Out-of-Distribution Detection and Policy Adaptation Robust control of markov decision processes with uncertain transition matrices.Operations Research, 53(5):780–798, 2005
Reference 70
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b541c8e6-7200-49c9-8b4e-0ae5f1afcb13 · outbound
Safe Domain Randomization via Uncertainty-Aware Out-of-Distribution Detection and Policy Adaptation Safe offline reinforcement learning with feasibility-guided diffusion model
Reference 71
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c5de7a59-d088-434c-bc24-6fa26922873b · outbound
Safe Domain Randomization via Uncertainty-Aware Out-of-Distribution Detection and Policy Adaptation Enhancing efficiency of safe reinforcement learning via sample manipulation
Reference 72
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 46eeeb28-8e94-4f3c-80ea-c5952b29cf04 · outbound
Safe Domain Randomization via Uncertainty-Aware Out-of-Distribution Detection and Policy Adaptation Curriculum learning for reinforcement learning domains: A framework and survey.Journal of Machine Learning Research, 21(181):1–50, 2020
Reference 73
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 26abd769-ff08-447c-97a1-4a6cacf296d7 · outbound
Safe Domain Randomization via Uncertainty-Aware Out-of-Distribution Detection and Policy Adaptation Causally aligned curriculum learning
Reference 74
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 44dea0fa-db63-40bf-b85a-0142e4bb4af6 · outbound
Safe Domain Randomization via Uncertainty-Aware Out-of-Distribution Detection and Policy Adaptation Au- tomated curriculum learning for neural networks
Reference 75
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7f211236-c2b5-42f5-a24d-baed0cfe0a8c · outbound
Safe Domain Randomization via Uncertainty-Aware Out-of-Distribution Detection and Policy Adaptation MIT press, 2016
Reference 76
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0b4402a7-9aa9-4766-b59a-a16ba0b5de03 · outbound
Safe Domain Randomization via Uncertainty-Aware Out-of-Distribution Detection and Policy Adaptation Robust training with ensemble consensus
Reference 77
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 3cb9dd47-d8f6-4661-b75f-d620146f0dda · outbound
Safe Domain Randomization via Uncertainty-Aware Out-of-Distribution Detection and Policy Adaptation Simple and scalable predictive uncertainty estimation using deep ensembles
Reference 78
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 08ad4517-5379-48f5-a9bb-e9a489d45773 · outbound
Safe Domain Randomization via Uncertainty-Aware Out-of-Distribution Detection and Policy Adaptation Improving robustness and calibration in ensembles with diversity regularization
Reference 79
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 8e4c9760-98c5-4ea7-bce9-973c718bbc77 · outbound
Safe Domain Randomization via Uncertainty-Aware Out-of-Distribution Detection and Policy Adaptation Maximizing overall diversity for improved uncertainty estimates in deep ensembles.Proceedings of the AAAI Conference on Artificial Intelligence, 34(04):4264–4271, Apr
Reference 80
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 497e7924-326b-42e8-856f-d687b10c2d4b · outbound
Safe Domain Randomization via Uncertainty-Aware Out-of-Distribution Detection and Policy Adaptation Improving adversarial robustness via promoting ensemble diversity
Reference 81
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 98359dcb-0d9c-4d54-9794-f517139dcc7e · outbound
Safe Domain Randomization via Uncertainty-Aware Out-of-Distribution Detection and Policy Adaptation Webb, Henry W
Reference 82
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 710834d2-13ba-415a-b6bd-246eb2e3a249 · outbound
Safe Domain Randomization via Uncertainty-Aware Out-of-Distribution Detection and Policy Adaptation Ensemble of averages: Improv- ing model selection and boosting performance in domain generalization
Reference 83
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation e261abd3-e9c6-43c6-ac43-171884808c09 · outbound
Safe Domain Randomization via Uncertainty-Aware Out-of-Distribution Detection and Policy Adaptation Noise contrastive priors for functional uncertainty
Reference 84
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 7818270d-e5f4-46ad-8d86-1c2645b3c400 · outbound
Safe Domain Randomization via Uncertainty-Aware Out-of-Distribution Detection and Policy Adaptation Deep exploration via bootstrapped dqn.Advances in neural information processing systems, 29, 2016
Reference 85
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 684f6ff6-8c3c-4228-81b9-26cf183d0e63 · outbound
Safe Domain Randomization via Uncertainty-Aware Out-of-Distribution Detection and Policy Adaptation Sunrise: A simple unified framework for ensemble learning in deep reinforcement learning
Reference 86
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 80a9eb5b-5761-4b2e-85d1-a21f6b0d1547 · outbound
Safe Domain Randomization via Uncertainty-Aware Out-of-Distribution Detection and Policy Adaptation Accurate uncertainty estimation and decomposition in ensemble learning.Advances in Neural Information Processing Systems, 32, 2019
Reference 87
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 941e3176-5522-4b1d-aa8a-c91731d50bd6 · outbound
Safe Domain Randomization via Uncertainty-Aware Out-of-Distribution Detection and Policy Adaptation Disentangling Epistemic and Aleatoric Uncertainty in Reinforcement Learning
Reference 88
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 42e34f71-1de2-4658-8ed3-6accdde04c70 · outbound
Safe Domain Randomization via Uncertainty-Aware Out-of-Distribution Detection and Policy Adaptation Kingma and Jimmy Ba
Reference 89
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8901a11e-c7fc-4d12-85bd-d8f69a797d87 · outbound
Safe Domain Randomization via Uncertainty-Aware Out-of-Distribution Detection and Policy Adaptation Proximal Policy Optimization Algorithms
Reference 90
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1b431425-3d5f-451c-b414-ff3fe25fd0c5 · outbound
Safe Domain Randomization via Uncertainty-Aware Out-of-Distribution Detection and Policy Adaptation |R(s, a)−R(s′, a′)| ≤LR (s, a)−(s′, a′) ,∀(s, a),(s ′, a′)∈S×A,(19) and satisfies|R(s, a)| ≤Rmax
Reference 91
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation e0c46e7f-3d9e-44d3-a648-8c53cd3bca6c · outbound
Safe Domain Randomization via Uncertainty-Aware Out-of-Distribution Detection and Policy Adaptation We acknowledge that real-world contact dynamics can violate global Lipschitz continuity
Reference 92
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation dd71c4ac-e2c6-459b-bc1e-a08fa2317c59 · outbound
Safe Domain Randomization via Uncertainty-Aware Out-of-Distribution Detection and Policy Adaptation distance
Reference 93
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation b8dd2bd5-6ccc-44d8-8534-bf3960160fc3 · outbound
Safe Domain Randomization via Uncertainty-Aware Out-of-Distribution Detection and Policy Adaptation • Depending on the country in which research is conducted, IRB approval (or equivalent) may be required for any human subjects research
Reference 94
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 574d3609-1717-4005-9922-f229e73f6ee2 · inbound
Toward Hardware-Agnostic Quadrupedal World Models via Morphology Conditioning Safe Domain Randomization via Uncertainty-Aware Out-of-Distribution Detection and Policy Adaptation
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 3d2a5ce3-47b7-41ea-91ad-ee34b05327d8 · inbound
Drift Q-Learning Safe Domain Randomization via Uncertainty-Aware Out-of-Distribution Detection and Policy Adaptation
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.