Pith. sign in

Paper Citation Record · LEDGER

DIAL: Distribution-Informed Adaptive Learning of Multi-Task Constraints for Safety-Critical Systems

As of 21 August 2026, this Paper Citation Record lists 63 of 63 outbound references and 0 inbound Pith citation observations for arXiv:2501.18086.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2501.18086 v1

Coverage vector

measured 63 of 63 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-10T00:48:32.753805Z

measured 63 of 63 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-21T06:32:19.484+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

63 of 63 outbound references displayed

  • verified exact1
  • verified fuzzy40
  • unresolved22
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 4daa092c-1248-4526-ba6a-f606389e8e78 · outbound

This paper cites Cic: Contrastive intrinsic control for unsupervised skill discovery,.

DIAL: Distribution-Informed Adaptive Learning of Multi-Task Constraints for Safety-Critical Systems Cic: Contrastive intrinsic control for unsupervised skill discovery,

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T00:48:35.028770Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-10T00:48:31.450693Z digest=sha256:9911fcd10d157e34488776f9f2e324189485aa458f26d2d5fc4be620243dd318

Observation 0342578d-af3c-4c21-9f96-9c7646e99500 · outbound

This paper cites Urlb: Unsupervised reinforcement learning benchmark,.

DIAL: Distribution-Informed Adaptive Learning of Multi-Task Constraints for Safety-Critical Systems Urlb: Unsupervised reinforcement learning benchmark,

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T00:48:35.012998Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-10T00:48:31.530551Z digest=sha256:f9ca7fe43dd6900bdd2a721195226b092fcf7d8a38c4b744db1ebf541af916c0

Observation 1df158d5-de82-4297-b5d9-14dfdad06d82 · outbound

This paper cites Unsupervised reinforcement learning in multiple environments,.

DIAL: Distribution-Informed Adaptive Learning of Multi-Task Constraints for Safety-Critical Systems Unsupervised reinforcement learning in multiple environments,

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T00:48:34.996288Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-10T00:48:31.583130Z digest=sha256:d764c60ad49ce322a8541fa7f2d911c5fd5046e2c72f337513d1a869e135a7c8

Observation 38bc8abc-7070-40ed-826c-43b1fb186ae6 · outbound

This paper cites an unresolved cited work.

DIAL: Distribution-Informed Adaptive Learning of Multi-Task Constraints for Safety-Critical Systems Unresolved cited work

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-10T00:48:31.635698Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T00:48:31.635698Z digest=sha256:d2ae18a96c86a144011ec71a1263d87903e4edcedcedd09aec55eebe65feda6d

Observation 7097344a-004f-47af-894a-e83ea9ccc664 · outbound

This paper cites Efficient training of artificial neural networks for autonomous navigation,.

DIAL: Distribution-Informed Adaptive Learning of Multi-Task Constraints for Safety-Critical Systems Efficient training of artificial neural networks for autonomous navigation,

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-10T00:48:31.675493Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T00:48:31.675493Z digest=sha256:d7657e192d64758a46bd09edbce3fa69481f044b18badd7e48939adc3714ff11

Observation 337df680-0a43-4008-bac2-0b5b3d6ee255 · outbound

This paper cites Learning from demonstration,.

DIAL: Distribution-Informed Adaptive Learning of Multi-Task Constraints for Safety-Critical Systems Learning from demonstration,

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T00:48:34.960046Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-10T00:48:31.683014Z digest=sha256:01dea4c65c1d5e540032684bea546516199fcb8f6b6ca63dde62fcda57dea514

Observation 1041deaa-0f38-48c6-9987-a05dfc64655e · outbound

This paper cites A survey of robot learning from demonstration,.

DIAL: Distribution-Informed Adaptive Learning of Multi-Task Constraints for Safety-Critical Systems A survey of robot learning from demonstration,

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-10T00:48:31.701992Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T00:48:31.701992Z digest=sha256:5ff1ebd7efbec42c5d08f430d430f29dd3109499399a8a045fa639ba52fd20fe

Observation 4c9b351d-e9e0-410e-88f9-1516913b0c8b · outbound

This paper cites Model-agnostic meta-learning for fast adaptation of deep networks,.

DIAL: Distribution-Informed Adaptive Learning of Multi-Task Constraints for Safety-Critical Systems Model-agnostic meta-learning for fast adaptation of deep networks,

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-10T00:48:31.724809Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T00:48:31.724809Z digest=sha256:34655f0451f1472ee8fab54b51ff3f926bc98e9dc146aa28e3da4777850e95ca

Observation a47ba89b-d833-4bd2-8b86-36926039bfeb · outbound

This paper cites Never give up: Learning directed exploration strategies,.

DIAL: Distribution-Informed Adaptive Learning of Multi-Task Constraints for Safety-Critical Systems Never give up: Learning directed exploration strategies,

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T00:48:34.924402Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-10T00:48:31.729354Z digest=sha256:b7adf7944f064cdd42c9b736c2470175249fad6953062c2720a5aa8e1ff643d5

Observation 776e3556-3fef-466d-9a36-ad0cfb2870e6 · outbound

This paper cites Novelty search in repre- sentational space for sample efficient exploration,.

DIAL: Distribution-Informed Adaptive Learning of Multi-Task Constraints for Safety-Critical Systems Novelty search in repre- sentational space for sample efficient exploration,

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T00:48:34.890509Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-10T00:48:31.734358Z digest=sha256:bab1b557eefa3daedc61900c91085f23287b2c248c82b0e5ee3ab007cdf4e98a

Observation e59bb6a3-7c2a-4cd5-b2ba-847ddb84b818 · outbound

This paper cites State entropy maximization with random encoders for efficient exploration,.

DIAL: Distribution-Informed Adaptive Learning of Multi-Task Constraints for Safety-Critical Systems State entropy maximization with random encoders for efficient exploration,

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-10T00:48:31.739472Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T00:48:31.739472Z digest=sha256:df542990f9bdcf905d3f3a16d6cb9b167959d3bfafa34123ff3230774597a043

Observation b502e8d5-6717-4515-856f-c542a6985f9f · outbound

This paper cites A comprehensive survey on safe reinforce- ment learning,.

DIAL: Distribution-Informed Adaptive Learning of Multi-Task Constraints for Safety-Critical Systems A comprehensive survey on safe reinforce- ment learning,

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-10T00:48:31.744381Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T00:48:31.744381Z digest=sha256:9f351a39b43e1807f83439023a67431f2934e5b5b51abebfb1392b498f798b78

Observation e73112fe-69e6-475e-83e4-dba79ea23a0e · outbound

This paper cites Learning to run a power network challenge for training topology controllers,.

DIAL: Distribution-Informed Adaptive Learning of Multi-Task Constraints for Safety-Critical Systems Learning to run a power network challenge for training topology controllers,

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T00:48:34.725677Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-10T00:48:31.749156Z digest=sha256:0ef8dcc950122832faafbe1cf770d979779cb61def8280fd7e714e05ca8c1d0d

Observation 096df9b7-499f-45ef-8b43-912f17e834c4 · outbound

This paper cites Challenges of real-world reinforcement learning: definitions, benchmarks and analysis,.

DIAL: Distribution-Informed Adaptive Learning of Multi-Task Constraints for Safety-Critical Systems Challenges of real-world reinforcement learning: definitions, benchmarks and analysis,

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T00:48:34.580303Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-10T00:48:31.754294Z digest=sha256:71840986226d84fea4182ed46524f30cceec12b55720d8aae1de5877391a955f

Observation 2c372663-fc1a-447c-ac08-4a5f29d452ef · outbound

This paper cites Constrained policy optimization,.

DIAL: Distribution-Informed Adaptive Learning of Multi-Task Constraints for Safety-Critical Systems Constrained policy optimization,

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-10T00:48:31.759043Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T00:48:31.759043Z digest=sha256:cb9aab10165bcbbe36ba47b5d8dc40f5de5699ce1a3b81e3ea04cda8be30ae43

Observation 2a4e7517-7172-4e4a-9d66-76d9eb1e4c36 · outbound

This paper cites Safe reinforcement learning in constrained markov decision processes,.

DIAL: Distribution-Informed Adaptive Learning of Multi-Task Constraints for Safety-Critical Systems Safe reinforcement learning in constrained markov decision processes,

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T00:48:34.515932Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-10T00:48:31.763786Z digest=sha256:3980fe9666b1a1133e4cf862dbaf962b1e4ab9338de7009b4146b4d769d9ed49

Observation 16eaab49-7206-4ea8-a3f6-aa7101110125 · outbound

This paper cites Density constrained reinforcement learning,.

DIAL: Distribution-Informed Adaptive Learning of Multi-Task Constraints for Safety-Critical Systems Density constrained reinforcement learning,

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T00:48:34.501166Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-10T00:48:31.769142Z digest=sha256:084274d1c3067967e93fe2a3329b0ad4778681afb874ad313960ab672f6a9793

Observation 5f59d091-fdf1-4802-85d4-e46607fa6fa0 · outbound

This paper cites Cem: Constrained entropy maximization for task-agnostic safe exploration,.

DIAL: Distribution-Informed Adaptive Learning of Multi-Task Constraints for Safety-Critical Systems Cem: Constrained entropy maximization for task-agnostic safe exploration,

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T00:48:34.486541Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-10T00:48:31.774386Z digest=sha256:7ffb64ec4bcc1d6aa9ae189f1214cf929b6a8ad6140bb1303d172ccdd468382e

Observation a1030c76-5be3-429f-8d8d-a443745378ac · outbound

This paper cites Learning constraints from demon- strations,.

DIAL: Distribution-Informed Adaptive Learning of Multi-Task Constraints for Safety-Critical Systems Learning constraints from demon- strations,

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T00:48:34.470519Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-10T00:48:31.779376Z digest=sha256:012c9c001cafcc0262656748993d019fe91e8fa92b9c08cd72fa6e531c335721

Observation 1f3df500-4e79-4011-9b80-cd7faaedc697 · outbound

This paper cites Inverse constrained re- inforcement learning,.

DIAL: Distribution-Informed Adaptive Learning of Multi-Task Constraints for Safety-Critical Systems Inverse constrained re- inforcement learning,

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T00:48:34.454623Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-10T00:48:31.852221Z digest=sha256:6ad65e355ccbbabf895c0a8cc93a5d6a1957f1db1139fab9cacb153d8e594609

Observation 1a9f6cfc-0b33-4415-ad61-d3857dffed08 · outbound

This paper cites Learning shared safety constraints from multi-task demonstrations,.

DIAL: Distribution-Informed Adaptive Learning of Multi-Task Constraints for Safety-Critical Systems Learning shared safety constraints from multi-task demonstrations,

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T00:48:34.438443Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-10T00:48:31.909696Z digest=sha256:4c90a1fba5347a7a1648089f4296c68c0000e65718b297ddf40c69add5b1ac45

Observation 19a93881-4d3d-465f-a7f1-0cc4bebc9677 · outbound

This paper cites Conditional value-at-risk for elliptical distributions,.

DIAL: Distribution-Informed Adaptive Learning of Multi-Task Constraints for Safety-Critical Systems Conditional value-at-risk for elliptical distributions,

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T00:48:34.422642Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-10T00:48:31.993829Z digest=sha256:496f7b8ba33df88a805e0c2659af53f4b4d4cecc8044afd76f4d4a1678013bdf

Observation f8c842f9-f593-4935-8117-231d643b78ab · outbound

This paper cites Worst cases policy gradients,.

DIAL: Distribution-Informed Adaptive Learning of Multi-Task Constraints for Safety-Critical Systems Worst cases policy gradients,

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T00:48:34.367155Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-10T00:48:32.117853Z digest=sha256:ff3e7adaee39b70e7ada8e331220000f0457e14ba7da23912f9864dcc72c14c7

Observation ee9d3446-804a-47ff-ac1a-70764afbcd6c · outbound

This paper cites Wcsac: Worst-case soft actor critic for safety-constrained reinforcement learn- ing,.

DIAL: Distribution-Informed Adaptive Learning of Multi-Task Constraints for Safety-Critical Systems Wcsac: Worst-case soft actor critic for safety-constrained reinforcement learn- ing,

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T00:48:34.176402Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-10T00:48:32.122955Z digest=sha256:770a993a7a35b5086a73c6cc5eea86c4b204e91a54a5b75fd69a4948e1c4d537

Observation 49f5489f-2fa0-4948-bbc2-5e4e295899ce · outbound

This paper cites Task-agnostic exploration in reinforce- ment learning,.

DIAL: Distribution-Informed Adaptive Learning of Multi-Task Constraints for Safety-Critical Systems Task-agnostic exploration in reinforce- ment learning,

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T00:48:34.088809Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-10T00:48:32.127213Z digest=sha256:fc5ec146b7ef8a891b1ec98140d7c0299643758cefe021fa93939b9cbf234d02

Observation 2396b1d5-c093-4948-aa9b-1f8e51078b4e · outbound

This paper cites Learning safety constraints from demonstrations with unknown re- wards,.

DIAL: Distribution-Informed Adaptive Learning of Multi-Task Constraints for Safety-Critical Systems Learning safety constraints from demonstrations with unknown re- wards,

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T00:48:34.073630Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-10T00:48:32.131439Z digest=sha256:44ed30c161f4b4cee1bbd5104cb8e52de3bd7f7251c3934d7944b1b2b1658a05

Observation f9918328-d66d-430b-b5ea-dd509971e848 · outbound

This paper cites Train hard, fight easy: Robust meta reinforcement learning,.

DIAL: Distribution-Informed Adaptive Learning of Multi-Task Constraints for Safety-Critical Systems Train hard, fight easy: Robust meta reinforcement learning,

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T00:48:34.057044Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-10T00:48:32.135777Z digest=sha256:e8a7a03a87c8cdcd38feb2b68c010409de50a4c7073ad252535085a07ef0ff82

Observation 0cabb9a8-e59f-45a7-8ccd-03f3d6dcfbc7 · outbound

This paper cites Incentivizing Exploration In Reinforcement Learning With Deep Predictive Models.

DIAL: Distribution-Informed Adaptive Learning of Multi-Task Constraints for Safety-Critical Systems Incentivizing Exploration In Reinforcement Learning With Deep Predictive Models

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-10T00:48:32.140574Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T00:48:32.140574Z digest=sha256:e815b9ef4a9785cedcae0df0fe7ac94b4870e900390a8736d49a3d82855a6153

Observation f0b238c4-f198-4e9f-ae0e-85ed1b072020 · outbound

This paper cites Reward-free exploration for reinforcement learning,.

DIAL: Distribution-Informed Adaptive Learning of Multi-Task Constraints for Safety-Critical Systems Reward-free exploration for reinforcement learning,

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T00:48:34.039793Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-10T00:48:32.146002Z digest=sha256:9fe59d78c00806c52e58aacd0e0c0c78147eec03a0f6ed21eb144e3a18dc2b8f

Observation a4fd1942-f300-4043-a1ec-0499e23eb5dc · outbound

This paper cites Provably efficient maximum entropy exploration,.

DIAL: Distribution-Informed Adaptive Learning of Multi-Task Constraints for Safety-Critical Systems Provably efficient maximum entropy exploration,

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T00:48:34.022731Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-10T00:48:32.150961Z digest=sha256:12807f22308790dc587dcc23c8ebe8f293bc9e87e60acd4da45111734fe70a81

Observation 46912794-0df4-4cfd-b9fc-38c387092c1b · outbound

This paper cites Efficient Exploration via State Marginal Matching.

DIAL: Distribution-Informed Adaptive Learning of Multi-Task Constraints for Safety-Critical Systems Efficient Exploration via State Marginal Matching

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-10T00:48:32.155017Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T00:48:32.155017Z digest=sha256:b4b24fc6db69ffcc9309f4583b15251d851ea2f132eded7e3a9002ba19d62e05

Observation 501e5005-bd15-45ff-98e7-714bcc3fe0c9 · outbound

This paper cites Constrained cross-entropy method for safe reinforcement learning,.

DIAL: Distribution-Informed Adaptive Learning of Multi-Task Constraints for Safety-Critical Systems Constrained cross-entropy method for safe reinforcement learning,

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T00:48:34.005886Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-10T00:48:32.160225Z digest=sha256:5e6f7e1d9bb1476aa3a66871e64c12cd0596e5f4cacc0f07aca9d6e0228a947e

Observation 8b45eaf0-623c-4d42-b538-8b8df7a11f90 · outbound

This paper cites Learning to fly,.

DIAL: Distribution-Informed Adaptive Learning of Multi-Task Constraints for Safety-Critical Systems Learning to fly,

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T00:48:33.990268Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-10T00:48:32.165376Z digest=sha256:1b4646fca8a8b5eba447fa563c98469a64019a40d299aec6d855f43835360cf3

Observation 936e9e12-8eab-4a6e-ba98-ed91ee1054d6 · outbound

This paper cites an unresolved cited work.

DIAL: Distribution-Informed Adaptive Learning of Multi-Task Constraints for Safety-Critical Systems Unresolved cited work

Reference 34

Resolution
unresolved
raw_fallback, observed 2026-08-10T00:48:33.973898Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-10T00:48:32.169801Z digest=sha256:05d54af10725c50cfc5c10412d474a1d4ad29fe4408be2b040fd8c57f2fb0fc0

Observation 3afd8016-a712-4b1d-8e5d-b78f1278f0f6 · outbound

This paper cites Apprenticeship learning via inverse rein- forcement learning,.

DIAL: Distribution-Informed Adaptive Learning of Multi-Task Constraints for Safety-Critical Systems Apprenticeship learning via inverse rein- forcement learning,

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-10T00:48:32.173834Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T00:48:32.173834Z digest=sha256:c9e72eb7d5741a4393bb3c763952db1562b064b3f2f43892b2c61f6168578b90

Observation b225c898-29e6-4348-a94b-32eaa65f2a24 · outbound

This paper cites Generative adversarial imitation learning,.

DIAL: Distribution-Informed Adaptive Learning of Multi-Task Constraints for Safety-Critical Systems Generative adversarial imitation learning,

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-10T00:48:32.178103Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T00:48:32.178103Z digest=sha256:61b17e10eeb725c75e6098269d18355b73a733450f3dfc571a408019eadc82b9

Observation 75b8c992-6b1e-4709-a1be-9e0eeed87ed5 · outbound

This paper cites Learning robust rewards with adverserial inverse reinforcement learning,.

DIAL: Distribution-Informed Adaptive Learning of Multi-Task Constraints for Safety-Critical Systems Learning robust rewards with adverserial inverse reinforcement learning,

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T00:48:33.939014Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-10T00:48:32.182217Z digest=sha256:fbee25741da750ea841265953e3210f7391657b8a0c7bb6164052fbe697d9a8c

Observation fa645a3a-0e85-4e29-8d65-0e5fea072805 · outbound

This paper cites Iq-learn: Inverse soft-q learning for imitation,.

DIAL: Distribution-Informed Adaptive Learning of Multi-Task Constraints for Safety-Critical Systems Iq-learn: Inverse soft-q learning for imitation,

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T00:48:33.923032Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-10T00:48:32.186768Z digest=sha256:0ff4240348272f8a6ff3364e1b68a2ceb3c4ed219e4473442e92b482f371d648

Observation 1816aead-b213-472b-a638-1e29af456356 · outbound

This paper cites Maximum entropy inverse reinforcement learning.

DIAL: Distribution-Informed Adaptive Learning of Multi-Task Constraints for Safety-Critical Systems Maximum entropy inverse reinforcement learning

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-10T00:48:32.191737Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T00:48:32.191737Z digest=sha256:409c3e7d5319995299dc643d0a3c1b007c4841ed4d77d516a1e8d5706652e88b

Observation 96faf043-56f6-454f-a84d-9e76eaea414d · outbound

This paper cites Bridging the gap between imitation learning and inverse reinforcement learning,.

DIAL: Distribution-Informed Adaptive Learning of Multi-Task Constraints for Safety-Critical Systems Bridging the gap between imitation learning and inverse reinforcement learning,

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T00:48:33.813002Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-10T00:48:32.196399Z digest=sha256:a490518bda23844ea09bb5c2a30b718964863c58f9e9fcdfdbc01ad8edd6dcc4

Observation f68e6c6e-c5ea-496a-8125-565e9aa65e62 · outbound

This paper cites A Connection between Generative Adversarial Networks, Inverse Reinforcement Learning, and Energy-Based Models.

DIAL: Distribution-Informed Adaptive Learning of Multi-Task Constraints for Safety-Critical Systems A Connection between Generative Adversarial Networks, Inverse Reinforcement Learning, and Energy-Based Models

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-10T00:48:32.201198Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T00:48:32.201198Z digest=sha256:8bff3e96eb1503d40c5967a2a4a48e5acff4b18570d1800e7a7a19435282b38e

Observation 80cb5e59-5b8b-49eb-86b1-83b401e1e587 · outbound

This paper cites Maximum Likelihood Constraint Inference for Inverse Reinforcement Learning.

DIAL: Distribution-Informed Adaptive Learning of Multi-Task Constraints for Safety-Critical Systems Maximum Likelihood Constraint Inference for Inverse Reinforcement Learning

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-10T00:48:32.206554Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T00:48:32.206554Z digest=sha256:6901d1e33d5dbe452553560ae950608b235966a532792cac0d540308c4e06f95

Observation fd1ded94-c38f-457b-bc72-e9148432ad45 · outbound

This paper cites Learning constraints from demon- strations with grid and parametric representations,.

DIAL: Distribution-Informed Adaptive Learning of Multi-Task Constraints for Safety-Critical Systems Learning constraints from demon- strations with grid and parametric representations,

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T00:48:33.715976Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-10T00:48:32.211728Z digest=sha256:efd4ca5b5177b35d3520fb5667cf7935ff1fe4ed0f1bdb40a55b490ce1960b40

Observation 573c1ce4-7b3d-4253-b4f9-f626d4ed3881 · outbound

This paper cites Gaussian process constraint learning for scalable chance-constrained motion planning from demon- strations,.

DIAL: Distribution-Informed Adaptive Learning of Multi-Task Constraints for Safety-Critical Systems Gaussian process constraint learning for scalable chance-constrained motion planning from demon- strations,

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T00:48:33.596229Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-10T00:48:32.216974Z digest=sha256:f02029436c08c7c554b9b9cb28c967f59e45ef13ed889ab4dbf160435188d250

Observation 1b5e8fdf-dd69-4838-92cb-de460f28a0b2 · outbound

This paper cites Learning soft constraints from constrained expert demonstrations,.

DIAL: Distribution-Informed Adaptive Learning of Multi-Task Constraints for Safety-Critical Systems Learning soft constraints from constrained expert demonstrations,

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T00:48:33.580418Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-10T00:48:32.222108Z digest=sha256:ddcff86a328606c6a2ee2ce3367aff8b639eb3deeec49ca84e74cd1a206d556c

Observation 1423062b-be2c-489e-8bdd-aea74a493fd5 · outbound

This paper cites Uncertainty-aware constraint inference in inverse constrained reinforcement learning,.

DIAL: Distribution-Informed Adaptive Learning of Multi-Task Constraints for Safety-Critical Systems Uncertainty-aware constraint inference in inverse constrained reinforcement learning,

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T00:48:33.564589Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-10T00:48:32.269252Z digest=sha256:a68968b4bb039d4df3a90b6e690c5fbb4f7bb361da7e268d2b75630e204977b0

Observation b2a25c70-9fb6-4c3b-8337-364794e6982c · outbound

This paper cites Confidence Aware Inverse Constrained Reinforcement Learning.

DIAL: Distribution-Informed Adaptive Learning of Multi-Task Constraints for Safety-Critical Systems Confidence Aware Inverse Constrained Reinforcement Learning

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-10T00:48:32.306249Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T00:48:32.306249Z digest=sha256:882ba3366b992be74775f5b9250ea9e7869e9d8e000308b1bcff533d24e297f8

Observation cc73d1fd-dd46-48ca-829f-9081f2c4c89f · outbound

This paper cites Inverse constraint learning and gen- eralization by transferable reward decomposition,.

DIAL: Distribution-Informed Adaptive Learning of Multi-Task Constraints for Safety-Critical Systems Inverse constraint learning and gen- eralization by transferable reward decomposition,

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T00:48:33.548668Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-10T00:48:32.342237Z digest=sha256:7f6ad5b5a9d21d60fbb991980213b3d6fbef44eb34b81d70ef32679cfdd9c049

Observation 27802c63-be74-4062-87c6-6c7945e75c81 · outbound

This paper cites Altman, Constrained Markov decision processes.

DIAL: Distribution-Informed Adaptive Learning of Multi-Task Constraints for Safety-Critical Systems Altman, Constrained Markov decision processes

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-10T00:48:32.455776Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T00:48:32.455776Z digest=sha256:9c091d5a57c8768bac0c7eaf0775f4efa7f8f35f9e3d2c2f8edaf93923352913

Observation db4b5a4d-393b-48bc-b55e-c00e0b9c856a · outbound

This paper cites Benchmarking Batch Deep Reinforcement Learning Algorithms.

DIAL: Distribution-Informed Adaptive Learning of Multi-Task Constraints for Safety-Critical Systems Benchmarking Batch Deep Reinforcement Learning Algorithms

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-10T00:48:32.560569Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T00:48:32.560569Z digest=sha256:c14a1a622a1b739a8c61127f6fd955d58ceaf2ad37a84305ba4a452d7aba807f

Observation 21b481c7-6aa3-4a04-ab70-bd745ffb3136 · outbound

This paper cites Approximate Robust Control of Uncertain Dynamical Systems.

DIAL: Distribution-Informed Adaptive Learning of Multi-Task Constraints for Safety-Critical Systems Approximate Robust Control of Uncertain Dynamical Systems

Reference 51

Resolution
verified exact
local_arxiv, observed 2026-08-10T00:48:32.854221Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-10T00:48:32.651094Z digest=sha256:47981bb74e4dd53b888754251b1c3652726aa426620bf1a03f04da703d645f44

Observation c3efdc1a-2d29-4ed2-9928-c5a8d085c1e7 · outbound

This paper cites Bayesian methods for constraint inference in reinforcement learning,.

DIAL: Distribution-Informed Adaptive Learning of Multi-Task Constraints for Safety-Critical Systems Bayesian methods for constraint inference in reinforcement learning,

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T00:48:33.522112Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-10T00:48:32.669783Z digest=sha256:eda052eaa3aaa62c3fa755c6042159d34045f937aa132a26084bad95f3758ce2

Observation 2a5dbb09-8942-4ec3-8318-6551f0845769 · outbound

This paper cites Risk-sensitive inverse reinforcement learning via coherent risk models.

DIAL: Distribution-Informed Adaptive Learning of Multi-Task Constraints for Safety-Critical Systems Risk-sensitive inverse reinforcement learning via coherent risk models

Reference 53

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T00:48:33.505778Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-10T00:48:32.684075Z digest=sha256:e14c6ce5368c440d392194b40cda1b4c3358ff46f1c80f10de6939accd44c228

Observation a0cbd5af-f9a4-4862-814c-e8e6f8d22f31 · outbound

This paper cites Benchmarking constraint inference in inverse reinforcement learning,.

DIAL: Distribution-Informed Adaptive Learning of Multi-Task Constraints for Safety-Critical Systems Benchmarking constraint inference in inverse reinforcement learning,

Reference 54

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T00:48:33.489286Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-10T00:48:32.690639Z digest=sha256:847f175c2f3d261def7cbc19c5a43ddd677595bef41f2dde19ba536ab18dc680

Observation 7150cc6a-b9b4-4b74-9a5e-8176cd2838e0 · outbound

This paper cites Near- est neighbor estimates of entropy,.

DIAL: Distribution-Informed Adaptive Learning of Multi-Task Constraints for Safety-Critical Systems Near- est neighbor estimates of entropy,

Reference 55

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T00:48:33.469520Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-10T00:48:32.701007Z digest=sha256:a38d5359693fe1ef7edea7ee237d52312c0397095daba38a1a5f263d9596c0bb

Observation c899864f-9de5-4213-a0f4-ef36e355e6e6 · outbound

This paper cites Particle based probability density fusion with differential shannon entropy criterion,.

DIAL: Distribution-Informed Adaptive Learning of Multi-Task Constraints for Safety-Critical Systems Particle based probability density fusion with differential shannon entropy criterion,

Reference 56

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T00:48:33.429642Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-10T00:48:32.709316Z digest=sha256:d8a5c78f04671bf68cf2db72026402322e3dad5564a5958b4c70bf667c012092

Observation 1da520f0-4726-42aa-bf2a-1a6e291dd466 · outbound

This paper cites Task-agnostic exploration via policy gradient of a non-parametric state entropy estimate,.

DIAL: Distribution-Informed Adaptive Learning of Multi-Task Constraints for Safety-Critical Systems Task-agnostic exploration via policy gradient of a non-parametric state entropy estimate,

Reference 57

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T00:48:33.374367Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-10T00:48:32.721815Z digest=sha256:73f507ec231abf61c2b0d69c32aa351624a5830877da4704fa2f79784fcb03e7

Observation 1eaa7da7-30cd-4d05-902e-ae9a0c322804 · outbound

This paper cites Reward Constrained Policy Optimization.

DIAL: Distribution-Informed Adaptive Learning of Multi-Task Constraints for Safety-Critical Systems Reward Constrained Policy Optimization

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-10T00:48:32.728968Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T00:48:32.728968Z digest=sha256:4b4794b0667248f602f29b2386ee34076f147bd387b2921d91ac9ec79ccba0ca

Observation bcf0f65e-4bce-42be-beee-e20cd5829ec7 · outbound

This paper cites An environment for autonomous driving decision- making,.

DIAL: Distribution-Informed Adaptive Learning of Multi-Task Constraints for Safety-Critical Systems An environment for autonomous driving decision- making,

Reference 59

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T00:48:33.235606Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-10T00:48:32.734110Z digest=sha256:6143ec0df5037df39a8ff0ba16713dc4b86778b3ebebbca7fed7f7ef8e93b359

Observation 1cc0b4a2-166f-4fe4-a896-a1de0e1827b2 · outbound

This paper cites OpenAI Gym.

DIAL: Distribution-Informed Adaptive Learning of Multi-Task Constraints for Safety-Critical Systems OpenAI Gym

Reference 60

Resolution
unresolved
no resolver link, observed 2026-08-10T00:48:32.739075Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T00:48:32.739075Z digest=sha256:ed35a5817759fdb530f17825ba72fe0c0065edd82ed5b5f7da7b47911b1ce3af

Observation a3c275c2-1d46-4086-94d7-f557ac3a4e64 · outbound

This paper cites Adam: A Method for Stochastic Optimization.

DIAL: Distribution-Informed Adaptive Learning of Multi-Task Constraints for Safety-Critical Systems Adam: A Method for Stochastic Optimization

Reference 61

Resolution
unresolved
no resolver link, observed 2026-08-10T00:48:32.744343Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T00:48:32.744343Z digest=sha256:ee51212aee23e37489638e90d29d4c94a411d802ebd91df928e8acbc638f71d5

Observation c6810a24-0bac-4704-8bd5-de1bad7855a6 · outbound

This paper cites Constrained differential optimization,.

DIAL: Distribution-Informed Adaptive Learning of Multi-Task Constraints for Safety-Critical Systems Constrained differential optimization,

Reference 62

Resolution
unresolved
no resolver link, observed 2026-08-10T00:48:32.749342Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T00:48:32.749342Z digest=sha256:b61a0a55969dcfa3903623f0758f0b1018509dd2772b0f6f073746ac5c8179ff

Observation a6d67648-b7bf-4c36-935d-b8118952f541 · outbound

This paper cites Controlled text gen- eration as continuous optimization with multiple constraints,.

DIAL: Distribution-Informed Adaptive Learning of Multi-Task Constraints for Safety-Critical Systems Controlled text gen- eration as continuous optimization with multiple constraints,

Reference 63

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T00:48:33.069704Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-10T00:48:32.753805Z digest=sha256:8e295d968bd9544deca7461b0362bfe35609099cf6f76ae5e9106157ee22106c

Pith citing papers

No inbound Pith citation observations are available.