Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-10T00:48:32.753805Z
Paper Citation Record · LEDGER
As of 21 August 2026, this Paper Citation Record lists 63 of 63 outbound references and 0 inbound Pith citation observations for arXiv:2501.18086.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-10T00:48:32.753805Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-20T06:33:59.587034+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
63 of 63 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 4daa092c-1248-4526-ba6a-f606389e8e78 · outbound
DIAL: Distribution-Informed Adaptive Learning of Multi-Task Constraints for Safety-Critical Systems Cic: Contrastive intrinsic control for unsupervised skill discovery,
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 0342578d-af3c-4c21-9f96-9c7646e99500 · outbound
DIAL: Distribution-Informed Adaptive Learning of Multi-Task Constraints for Safety-Critical Systems Urlb: Unsupervised reinforcement learning benchmark,
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 1df158d5-de82-4297-b5d9-14dfdad06d82 · outbound
DIAL: Distribution-Informed Adaptive Learning of Multi-Task Constraints for Safety-Critical Systems Unsupervised reinforcement learning in multiple environments,
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 38bc8abc-7070-40ed-826c-43b1fb186ae6 · outbound
DIAL: Distribution-Informed Adaptive Learning of Multi-Task Constraints for Safety-Critical Systems Unresolved cited work
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7097344a-004f-47af-894a-e83ea9ccc664 · outbound
DIAL: Distribution-Informed Adaptive Learning of Multi-Task Constraints for Safety-Critical Systems Efficient training of artificial neural networks for autonomous navigation,
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 337df680-0a43-4008-bac2-0b5b3d6ee255 · outbound
DIAL: Distribution-Informed Adaptive Learning of Multi-Task Constraints for Safety-Critical Systems Learning from demonstration,
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 1041deaa-0f38-48c6-9987-a05dfc64655e · outbound
DIAL: Distribution-Informed Adaptive Learning of Multi-Task Constraints for Safety-Critical Systems A survey of robot learning from demonstration,
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4c9b351d-e9e0-410e-88f9-1516913b0c8b · outbound
DIAL: Distribution-Informed Adaptive Learning of Multi-Task Constraints for Safety-Critical Systems Model-agnostic meta-learning for fast adaptation of deep networks,
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a47ba89b-d833-4bd2-8b86-36926039bfeb · outbound
DIAL: Distribution-Informed Adaptive Learning of Multi-Task Constraints for Safety-Critical Systems Never give up: Learning directed exploration strategies,
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 776e3556-3fef-466d-9a36-ad0cfb2870e6 · outbound
DIAL: Distribution-Informed Adaptive Learning of Multi-Task Constraints for Safety-Critical Systems Novelty search in repre- sentational space for sample efficient exploration,
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation e59bb6a3-7c2a-4cd5-b2ba-847ddb84b818 · outbound
DIAL: Distribution-Informed Adaptive Learning of Multi-Task Constraints for Safety-Critical Systems State entropy maximization with random encoders for efficient exploration,
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b502e8d5-6717-4515-856f-c542a6985f9f · outbound
DIAL: Distribution-Informed Adaptive Learning of Multi-Task Constraints for Safety-Critical Systems A comprehensive survey on safe reinforce- ment learning,
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e73112fe-69e6-475e-83e4-dba79ea23a0e · outbound
DIAL: Distribution-Informed Adaptive Learning of Multi-Task Constraints for Safety-Critical Systems Learning to run a power network challenge for training topology controllers,
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 096df9b7-499f-45ef-8b43-912f17e834c4 · outbound
DIAL: Distribution-Informed Adaptive Learning of Multi-Task Constraints for Safety-Critical Systems Challenges of real-world reinforcement learning: definitions, benchmarks and analysis,
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 2c372663-fc1a-447c-ac08-4a5f29d452ef · outbound
DIAL: Distribution-Informed Adaptive Learning of Multi-Task Constraints for Safety-Critical Systems Constrained policy optimization,
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2a4e7517-7172-4e4a-9d66-76d9eb1e4c36 · outbound
DIAL: Distribution-Informed Adaptive Learning of Multi-Task Constraints for Safety-Critical Systems Safe reinforcement learning in constrained markov decision processes,
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 16eaab49-7206-4ea8-a3f6-aa7101110125 · outbound
DIAL: Distribution-Informed Adaptive Learning of Multi-Task Constraints for Safety-Critical Systems Density constrained reinforcement learning,
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 5f59d091-fdf1-4802-85d4-e46607fa6fa0 · outbound
DIAL: Distribution-Informed Adaptive Learning of Multi-Task Constraints for Safety-Critical Systems Cem: Constrained entropy maximization for task-agnostic safe exploration,
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation a1030c76-5be3-429f-8d8d-a443745378ac · outbound
DIAL: Distribution-Informed Adaptive Learning of Multi-Task Constraints for Safety-Critical Systems Learning constraints from demon- strations,
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 1f3df500-4e79-4011-9b80-cd7faaedc697 · outbound
DIAL: Distribution-Informed Adaptive Learning of Multi-Task Constraints for Safety-Critical Systems Inverse constrained re- inforcement learning,
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 1a9f6cfc-0b33-4415-ad61-d3857dffed08 · outbound
DIAL: Distribution-Informed Adaptive Learning of Multi-Task Constraints for Safety-Critical Systems Learning shared safety constraints from multi-task demonstrations,
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 19a93881-4d3d-465f-a7f1-0cc4bebc9677 · outbound
DIAL: Distribution-Informed Adaptive Learning of Multi-Task Constraints for Safety-Critical Systems Conditional value-at-risk for elliptical distributions,
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation f8c842f9-f593-4935-8117-231d643b78ab · outbound
DIAL: Distribution-Informed Adaptive Learning of Multi-Task Constraints for Safety-Critical Systems Worst cases policy gradients,
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation ee9d3446-804a-47ff-ac1a-70764afbcd6c · outbound
DIAL: Distribution-Informed Adaptive Learning of Multi-Task Constraints for Safety-Critical Systems Wcsac: Worst-case soft actor critic for safety-constrained reinforcement learn- ing,
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 49f5489f-2fa0-4948-bbc2-5e4e295899ce · outbound
DIAL: Distribution-Informed Adaptive Learning of Multi-Task Constraints for Safety-Critical Systems Task-agnostic exploration in reinforce- ment learning,
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 2396b1d5-c093-4948-aa9b-1f8e51078b4e · outbound
DIAL: Distribution-Informed Adaptive Learning of Multi-Task Constraints for Safety-Critical Systems Learning safety constraints from demonstrations with unknown re- wards,
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation f9918328-d66d-430b-b5ea-dd509971e848 · outbound
DIAL: Distribution-Informed Adaptive Learning of Multi-Task Constraints for Safety-Critical Systems Train hard, fight easy: Robust meta reinforcement learning,
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 0cabb9a8-e59f-45a7-8ccd-03f3d6dcfbc7 · outbound
DIAL: Distribution-Informed Adaptive Learning of Multi-Task Constraints for Safety-Critical Systems Incentivizing Exploration In Reinforcement Learning With Deep Predictive Models
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f0b238c4-f198-4e9f-ae0e-85ed1b072020 · outbound
DIAL: Distribution-Informed Adaptive Learning of Multi-Task Constraints for Safety-Critical Systems Reward-free exploration for reinforcement learning,
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation a4fd1942-f300-4043-a1ec-0499e23eb5dc · outbound
DIAL: Distribution-Informed Adaptive Learning of Multi-Task Constraints for Safety-Critical Systems Provably efficient maximum entropy exploration,
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 46912794-0df4-4cfd-b9fc-38c387092c1b · outbound
DIAL: Distribution-Informed Adaptive Learning of Multi-Task Constraints for Safety-Critical Systems Efficient Exploration via State Marginal Matching
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 501e5005-bd15-45ff-98e7-714bcc3fe0c9 · outbound
DIAL: Distribution-Informed Adaptive Learning of Multi-Task Constraints for Safety-Critical Systems Constrained cross-entropy method for safe reinforcement learning,
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 8b45eaf0-623c-4d42-b538-8b8df7a11f90 · outbound
DIAL: Distribution-Informed Adaptive Learning of Multi-Task Constraints for Safety-Critical Systems Learning to fly,
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 936e9e12-8eab-4a6e-ba98-ed91ee1054d6 · outbound
DIAL: Distribution-Informed Adaptive Learning of Multi-Task Constraints for Safety-Critical Systems Unresolved cited work
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 3afd8016-a712-4b1d-8e5d-b78f1278f0f6 · outbound
DIAL: Distribution-Informed Adaptive Learning of Multi-Task Constraints for Safety-Critical Systems Apprenticeship learning via inverse rein- forcement learning,
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b225c898-29e6-4348-a94b-32eaa65f2a24 · outbound
DIAL: Distribution-Informed Adaptive Learning of Multi-Task Constraints for Safety-Critical Systems Generative adversarial imitation learning,
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 75b8c992-6b1e-4709-a1be-9e0eeed87ed5 · outbound
DIAL: Distribution-Informed Adaptive Learning of Multi-Task Constraints for Safety-Critical Systems Learning robust rewards with adverserial inverse reinforcement learning,
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation fa645a3a-0e85-4e29-8d65-0e5fea072805 · outbound
DIAL: Distribution-Informed Adaptive Learning of Multi-Task Constraints for Safety-Critical Systems Iq-learn: Inverse soft-q learning for imitation,
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 1816aead-b213-472b-a638-1e29af456356 · outbound
DIAL: Distribution-Informed Adaptive Learning of Multi-Task Constraints for Safety-Critical Systems Maximum entropy inverse reinforcement learning
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 96faf043-56f6-454f-a84d-9e76eaea414d · outbound
DIAL: Distribution-Informed Adaptive Learning of Multi-Task Constraints for Safety-Critical Systems Bridging the gap between imitation learning and inverse reinforcement learning,
Reference 40
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation f68e6c6e-c5ea-496a-8125-565e9aa65e62 · outbound
DIAL: Distribution-Informed Adaptive Learning of Multi-Task Constraints for Safety-Critical Systems A Connection between Generative Adversarial Networks, Inverse Reinforcement Learning, and Energy-Based Models
Reference 41
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 80cb5e59-5b8b-49eb-86b1-83b401e1e587 · outbound
DIAL: Distribution-Informed Adaptive Learning of Multi-Task Constraints for Safety-Critical Systems Maximum Likelihood Constraint Inference for Inverse Reinforcement Learning
Reference 42
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fd1ded94-c38f-457b-bc72-e9148432ad45 · outbound
DIAL: Distribution-Informed Adaptive Learning of Multi-Task Constraints for Safety-Critical Systems Learning constraints from demon- strations with grid and parametric representations,
Reference 43
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 573c1ce4-7b3d-4253-b4f9-f626d4ed3881 · outbound
DIAL: Distribution-Informed Adaptive Learning of Multi-Task Constraints for Safety-Critical Systems Gaussian process constraint learning for scalable chance-constrained motion planning from demon- strations,
Reference 44
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 1b5e8fdf-dd69-4838-92cb-de460f28a0b2 · outbound
DIAL: Distribution-Informed Adaptive Learning of Multi-Task Constraints for Safety-Critical Systems Learning soft constraints from constrained expert demonstrations,
Reference 45
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 1423062b-be2c-489e-8bdd-aea74a493fd5 · outbound
DIAL: Distribution-Informed Adaptive Learning of Multi-Task Constraints for Safety-Critical Systems Uncertainty-aware constraint inference in inverse constrained reinforcement learning,
Reference 46
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation b2a25c70-9fb6-4c3b-8337-364794e6982c · outbound
DIAL: Distribution-Informed Adaptive Learning of Multi-Task Constraints for Safety-Critical Systems Confidence Aware Inverse Constrained Reinforcement Learning
Reference 47
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cc73d1fd-dd46-48ca-829f-9081f2c4c89f · outbound
DIAL: Distribution-Informed Adaptive Learning of Multi-Task Constraints for Safety-Critical Systems Inverse constraint learning and gen- eralization by transferable reward decomposition,
Reference 48
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 27802c63-be74-4062-87c6-6c7945e75c81 · outbound
DIAL: Distribution-Informed Adaptive Learning of Multi-Task Constraints for Safety-Critical Systems Altman, Constrained Markov decision processes
Reference 49
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation db4b5a4d-393b-48bc-b55e-c00e0b9c856a · outbound
DIAL: Distribution-Informed Adaptive Learning of Multi-Task Constraints for Safety-Critical Systems Benchmarking Batch Deep Reinforcement Learning Algorithms
Reference 50
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 21b481c7-6aa3-4a04-ab70-bd745ffb3136 · outbound
DIAL: Distribution-Informed Adaptive Learning of Multi-Task Constraints for Safety-Critical Systems Approximate Robust Control of Uncertain Dynamical Systems
Reference 51
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation c3efdc1a-2d29-4ed2-9928-c5a8d085c1e7 · outbound
DIAL: Distribution-Informed Adaptive Learning of Multi-Task Constraints for Safety-Critical Systems Bayesian methods for constraint inference in reinforcement learning,
Reference 52
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 2a5dbb09-8942-4ec3-8318-6551f0845769 · outbound
DIAL: Distribution-Informed Adaptive Learning of Multi-Task Constraints for Safety-Critical Systems Risk-sensitive inverse reinforcement learning via coherent risk models
Reference 53
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation a0cbd5af-f9a4-4862-814c-e8e6f8d22f31 · outbound
DIAL: Distribution-Informed Adaptive Learning of Multi-Task Constraints for Safety-Critical Systems Benchmarking constraint inference in inverse reinforcement learning,
Reference 54
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 7150cc6a-b9b4-4b74-9a5e-8176cd2838e0 · outbound
DIAL: Distribution-Informed Adaptive Learning of Multi-Task Constraints for Safety-Critical Systems Near- est neighbor estimates of entropy,
Reference 55
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation c899864f-9de5-4213-a0f4-ef36e355e6e6 · outbound
DIAL: Distribution-Informed Adaptive Learning of Multi-Task Constraints for Safety-Critical Systems Particle based probability density fusion with differential shannon entropy criterion,
Reference 56
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 1da520f0-4726-42aa-bf2a-1a6e291dd466 · outbound
DIAL: Distribution-Informed Adaptive Learning of Multi-Task Constraints for Safety-Critical Systems Task-agnostic exploration via policy gradient of a non-parametric state entropy estimate,
Reference 57
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 1eaa7da7-30cd-4d05-902e-ae9a0c322804 · outbound
DIAL: Distribution-Informed Adaptive Learning of Multi-Task Constraints for Safety-Critical Systems Reward Constrained Policy Optimization
Reference 58
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bcf0f65e-4bce-42be-beee-e20cd5829ec7 · outbound
DIAL: Distribution-Informed Adaptive Learning of Multi-Task Constraints for Safety-Critical Systems An environment for autonomous driving decision- making,
Reference 59
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 1cc0b4a2-166f-4fe4-a896-a1de0e1827b2 · outbound
DIAL: Distribution-Informed Adaptive Learning of Multi-Task Constraints for Safety-Critical Systems OpenAI Gym
Reference 60
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a3c275c2-1d46-4086-94d7-f557ac3a4e64 · outbound
DIAL: Distribution-Informed Adaptive Learning of Multi-Task Constraints for Safety-Critical Systems Adam: A Method for Stochastic Optimization
Reference 61
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c6810a24-0bac-4704-8bd5-de1bad7855a6 · outbound
DIAL: Distribution-Informed Adaptive Learning of Multi-Task Constraints for Safety-Critical Systems Constrained differential optimization,
Reference 62
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a6d67648-b7bf-4c36-935d-b8118952f541 · outbound
DIAL: Distribution-Informed Adaptive Learning of Multi-Task Constraints for Safety-Critical Systems Controlled text gen- eration as continuous optimization with multiple constraints,
Reference 63
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
No inbound Pith citation observations are available.