Pith. sign in

Paper Citation Record · LEDGER

DIAL: Distribution-Informed Adaptive Learning of Multi-Task Constraints for Safety-Critical Systems

As of 21 August 2026, this Paper Citation Record lists 63 of 63 outbound references and 0 inbound Pith citation observations for arXiv:2501.18086.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2501.18086 v1

Coverage vector

measured 63 of 63 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-10T00:48:32.753805Z

measured 63 of 63 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-20T06:33:59.587034+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

63 of 63 outbound references displayed

  • verified exact1
  • verified fuzzy40
  • unresolved22
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 4daa092c-1248-4526-ba6a-f606389e8e78 · outbound

This paper cites Cic: Contrastive intrinsic control for unsupervised skill discovery,.

DIAL: Distribution-Informed Adaptive Learning of Multi-Task Constraints for Safety-Critical Systems Cic: Contrastive intrinsic control for unsupervised skill discovery,

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T00:48:35.028770Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-10T00:48:31.450693Z digest=sha256:8738b8b7f6cac0b051252a70132677a36c05e6f9bbfb4cfc3c8865dff04206da

Observation 0342578d-af3c-4c21-9f96-9c7646e99500 · outbound

This paper cites Urlb: Unsupervised reinforcement learning benchmark,.

DIAL: Distribution-Informed Adaptive Learning of Multi-Task Constraints for Safety-Critical Systems Urlb: Unsupervised reinforcement learning benchmark,

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T00:48:35.012998Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-10T00:48:31.530551Z digest=sha256:4760599de0f920a0e28ab90c8a293e754bcbb740842411a15e02a99460ba2f59

Observation 1df158d5-de82-4297-b5d9-14dfdad06d82 · outbound

This paper cites Unsupervised reinforcement learning in multiple environments,.

DIAL: Distribution-Informed Adaptive Learning of Multi-Task Constraints for Safety-Critical Systems Unsupervised reinforcement learning in multiple environments,

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T00:48:34.996288Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-10T00:48:31.583130Z digest=sha256:ff1832a126c091eafd0f7ad67f30a7b5e89d2a198216171c676457594f0fc4f7

Observation 38bc8abc-7070-40ed-826c-43b1fb186ae6 · outbound

This paper cites an unresolved cited work.

DIAL: Distribution-Informed Adaptive Learning of Multi-Task Constraints for Safety-Critical Systems Unresolved cited work

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-10T00:48:31.635698Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T00:48:31.635698Z digest=sha256:d2ae18a96c86a144011ec71a1263d87903e4edcedcedd09aec55eebe65feda6d

Observation 7097344a-004f-47af-894a-e83ea9ccc664 · outbound

This paper cites Efficient training of artificial neural networks for autonomous navigation,.

DIAL: Distribution-Informed Adaptive Learning of Multi-Task Constraints for Safety-Critical Systems Efficient training of artificial neural networks for autonomous navigation,

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-10T00:48:31.675493Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T00:48:31.675493Z digest=sha256:d7657e192d64758a46bd09edbce3fa69481f044b18badd7e48939adc3714ff11

Observation 337df680-0a43-4008-bac2-0b5b3d6ee255 · outbound

This paper cites Learning from demonstration,.

DIAL: Distribution-Informed Adaptive Learning of Multi-Task Constraints for Safety-Critical Systems Learning from demonstration,

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T00:48:34.960046Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-10T00:48:31.683014Z digest=sha256:0fe2abb9dfbaed272570ed1ef1351776401acebc155183ec0845dfd418d27533

Observation 1041deaa-0f38-48c6-9987-a05dfc64655e · outbound

This paper cites A survey of robot learning from demonstration,.

DIAL: Distribution-Informed Adaptive Learning of Multi-Task Constraints for Safety-Critical Systems A survey of robot learning from demonstration,

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-10T00:48:31.701992Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T00:48:31.701992Z digest=sha256:5ff1ebd7efbec42c5d08f430d430f29dd3109499399a8a045fa639ba52fd20fe

Observation 4c9b351d-e9e0-410e-88f9-1516913b0c8b · outbound

This paper cites Model-agnostic meta-learning for fast adaptation of deep networks,.

DIAL: Distribution-Informed Adaptive Learning of Multi-Task Constraints for Safety-Critical Systems Model-agnostic meta-learning for fast adaptation of deep networks,

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-10T00:48:31.724809Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T00:48:31.724809Z digest=sha256:34655f0451f1472ee8fab54b51ff3f926bc98e9dc146aa28e3da4777850e95ca

Observation a47ba89b-d833-4bd2-8b86-36926039bfeb · outbound

This paper cites Never give up: Learning directed exploration strategies,.

DIAL: Distribution-Informed Adaptive Learning of Multi-Task Constraints for Safety-Critical Systems Never give up: Learning directed exploration strategies,

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T00:48:34.924402Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-10T00:48:31.729354Z digest=sha256:bb04dc6c7deaae277e3baaa562594fc964d4c242489d35c7bddd4a8e7552db08

Observation 776e3556-3fef-466d-9a36-ad0cfb2870e6 · outbound

This paper cites Novelty search in repre- sentational space for sample efficient exploration,.

DIAL: Distribution-Informed Adaptive Learning of Multi-Task Constraints for Safety-Critical Systems Novelty search in repre- sentational space for sample efficient exploration,

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T00:48:34.890509Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-10T00:48:31.734358Z digest=sha256:06935b26a51ce3f9d7347d2a9490e8591f12d8d4bf787111080569f9b0d36b58

Observation e59bb6a3-7c2a-4cd5-b2ba-847ddb84b818 · outbound

This paper cites State entropy maximization with random encoders for efficient exploration,.

DIAL: Distribution-Informed Adaptive Learning of Multi-Task Constraints for Safety-Critical Systems State entropy maximization with random encoders for efficient exploration,

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-10T00:48:31.739472Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T00:48:31.739472Z digest=sha256:df542990f9bdcf905d3f3a16d6cb9b167959d3bfafa34123ff3230774597a043

Observation b502e8d5-6717-4515-856f-c542a6985f9f · outbound

This paper cites A comprehensive survey on safe reinforce- ment learning,.

DIAL: Distribution-Informed Adaptive Learning of Multi-Task Constraints for Safety-Critical Systems A comprehensive survey on safe reinforce- ment learning,

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-10T00:48:31.744381Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T00:48:31.744381Z digest=sha256:9f351a39b43e1807f83439023a67431f2934e5b5b51abebfb1392b498f798b78

Observation e73112fe-69e6-475e-83e4-dba79ea23a0e · outbound

This paper cites Learning to run a power network challenge for training topology controllers,.

DIAL: Distribution-Informed Adaptive Learning of Multi-Task Constraints for Safety-Critical Systems Learning to run a power network challenge for training topology controllers,

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T00:48:34.725677Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-10T00:48:31.749156Z digest=sha256:9e40a8cba31c2d98ddfb14fd8cdef69a473411bf52980ab52545612ff3d5e1cb

Observation 096df9b7-499f-45ef-8b43-912f17e834c4 · outbound

This paper cites Challenges of real-world reinforcement learning: definitions, benchmarks and analysis,.

DIAL: Distribution-Informed Adaptive Learning of Multi-Task Constraints for Safety-Critical Systems Challenges of real-world reinforcement learning: definitions, benchmarks and analysis,

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T00:48:34.580303Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-10T00:48:31.754294Z digest=sha256:b0458259ea86fddb6bd9affa88d392f1616ac3e995e73580a74ff2a83021ad34

Observation 2c372663-fc1a-447c-ac08-4a5f29d452ef · outbound

This paper cites Constrained policy optimization,.

DIAL: Distribution-Informed Adaptive Learning of Multi-Task Constraints for Safety-Critical Systems Constrained policy optimization,

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-10T00:48:31.759043Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T00:48:31.759043Z digest=sha256:cb9aab10165bcbbe36ba47b5d8dc40f5de5699ce1a3b81e3ea04cda8be30ae43

Observation 2a4e7517-7172-4e4a-9d66-76d9eb1e4c36 · outbound

This paper cites Safe reinforcement learning in constrained markov decision processes,.

DIAL: Distribution-Informed Adaptive Learning of Multi-Task Constraints for Safety-Critical Systems Safe reinforcement learning in constrained markov decision processes,

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T00:48:34.515932Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-10T00:48:31.763786Z digest=sha256:d7f5deaa4ff90ca16ae6c4963abf187396ca063f496e415bfd0120ebc97a6871

Observation 16eaab49-7206-4ea8-a3f6-aa7101110125 · outbound

This paper cites Density constrained reinforcement learning,.

DIAL: Distribution-Informed Adaptive Learning of Multi-Task Constraints for Safety-Critical Systems Density constrained reinforcement learning,

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T00:48:34.501166Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-10T00:48:31.769142Z digest=sha256:f0d6af7630704e08d1ca6f5f57a222d6faeec6beb31d5a4d724c83f8952a4648

Observation 5f59d091-fdf1-4802-85d4-e46607fa6fa0 · outbound

This paper cites Cem: Constrained entropy maximization for task-agnostic safe exploration,.

DIAL: Distribution-Informed Adaptive Learning of Multi-Task Constraints for Safety-Critical Systems Cem: Constrained entropy maximization for task-agnostic safe exploration,

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T00:48:34.486541Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-10T00:48:31.774386Z digest=sha256:24ff1f70253d952c9dcbf7c53692af6257f8a1ffa9c6259dc07da7e2cee55e23

Observation a1030c76-5be3-429f-8d8d-a443745378ac · outbound

This paper cites Learning constraints from demon- strations,.

DIAL: Distribution-Informed Adaptive Learning of Multi-Task Constraints for Safety-Critical Systems Learning constraints from demon- strations,

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T00:48:34.470519Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-10T00:48:31.779376Z digest=sha256:cd41ce18c77cbfef7fb4a3f7a670bfd85076a39b41c709c578a76fde549d4102

Observation 1f3df500-4e79-4011-9b80-cd7faaedc697 · outbound

This paper cites Inverse constrained re- inforcement learning,.

DIAL: Distribution-Informed Adaptive Learning of Multi-Task Constraints for Safety-Critical Systems Inverse constrained re- inforcement learning,

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T00:48:34.454623Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-10T00:48:31.852221Z digest=sha256:7028b5f89a49545e0a1f97ac1814335cbbf730ba8c62102ffe1ae2471c056f78

Observation 1a9f6cfc-0b33-4415-ad61-d3857dffed08 · outbound

This paper cites Learning shared safety constraints from multi-task demonstrations,.

DIAL: Distribution-Informed Adaptive Learning of Multi-Task Constraints for Safety-Critical Systems Learning shared safety constraints from multi-task demonstrations,

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T00:48:34.438443Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-10T00:48:31.909696Z digest=sha256:162b85b6f9acab6a59e0dc910e5d2b741859c64f8b731e70049bddd2468f5e41

Observation 19a93881-4d3d-465f-a7f1-0cc4bebc9677 · outbound

This paper cites Conditional value-at-risk for elliptical distributions,.

DIAL: Distribution-Informed Adaptive Learning of Multi-Task Constraints for Safety-Critical Systems Conditional value-at-risk for elliptical distributions,

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T00:48:34.422642Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-10T00:48:31.993829Z digest=sha256:9f239430fa80d928cc2e138f90db5369d20c1cb6ab9d9c31be045e67c877fd40

Observation f8c842f9-f593-4935-8117-231d643b78ab · outbound

This paper cites Worst cases policy gradients,.

DIAL: Distribution-Informed Adaptive Learning of Multi-Task Constraints for Safety-Critical Systems Worst cases policy gradients,

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T00:48:34.367155Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-10T00:48:32.117853Z digest=sha256:f1c7cecd980864b2fcc9863e5d45dc729b38040746355ccdb8341c42d5dad1a6

Observation ee9d3446-804a-47ff-ac1a-70764afbcd6c · outbound

This paper cites Wcsac: Worst-case soft actor critic for safety-constrained reinforcement learn- ing,.

DIAL: Distribution-Informed Adaptive Learning of Multi-Task Constraints for Safety-Critical Systems Wcsac: Worst-case soft actor critic for safety-constrained reinforcement learn- ing,

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T00:48:34.176402Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-10T00:48:32.122955Z digest=sha256:9f765055305c905d8790ee63a0f0d9ef032b8c311421aca63538d1c5824de88a

Observation 49f5489f-2fa0-4948-bbc2-5e4e295899ce · outbound

This paper cites Task-agnostic exploration in reinforce- ment learning,.

DIAL: Distribution-Informed Adaptive Learning of Multi-Task Constraints for Safety-Critical Systems Task-agnostic exploration in reinforce- ment learning,

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T00:48:34.088809Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-10T00:48:32.127213Z digest=sha256:4b589291f251cddc8154c616bbcc2be42b2fcdf1f49742d515050a7a4b5ebada

Observation 2396b1d5-c093-4948-aa9b-1f8e51078b4e · outbound

This paper cites Learning safety constraints from demonstrations with unknown re- wards,.

DIAL: Distribution-Informed Adaptive Learning of Multi-Task Constraints for Safety-Critical Systems Learning safety constraints from demonstrations with unknown re- wards,

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T00:48:34.073630Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-10T00:48:32.131439Z digest=sha256:80ea4510e549b37b435296b5059d1cc6d13bf27c9387628e152c9fb7e2c9c054

Observation f9918328-d66d-430b-b5ea-dd509971e848 · outbound

This paper cites Train hard, fight easy: Robust meta reinforcement learning,.

DIAL: Distribution-Informed Adaptive Learning of Multi-Task Constraints for Safety-Critical Systems Train hard, fight easy: Robust meta reinforcement learning,

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T00:48:34.057044Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-10T00:48:32.135777Z digest=sha256:592fc104d33707da824dd2fb8a54cce0baccac5774fe85315327a65abf7f1fa8

Observation 0cabb9a8-e59f-45a7-8ccd-03f3d6dcfbc7 · outbound

This paper cites Incentivizing Exploration In Reinforcement Learning With Deep Predictive Models.

DIAL: Distribution-Informed Adaptive Learning of Multi-Task Constraints for Safety-Critical Systems Incentivizing Exploration In Reinforcement Learning With Deep Predictive Models

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-10T00:48:32.140574Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T00:48:32.140574Z digest=sha256:e815b9ef4a9785cedcae0df0fe7ac94b4870e900390a8736d49a3d82855a6153

Observation f0b238c4-f198-4e9f-ae0e-85ed1b072020 · outbound

This paper cites Reward-free exploration for reinforcement learning,.

DIAL: Distribution-Informed Adaptive Learning of Multi-Task Constraints for Safety-Critical Systems Reward-free exploration for reinforcement learning,

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T00:48:34.039793Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-10T00:48:32.146002Z digest=sha256:554d9c0144f1d4f7c30dafe0a72f9b735240a5522ba1162aed70406dc2d65257

Observation a4fd1942-f300-4043-a1ec-0499e23eb5dc · outbound

This paper cites Provably efficient maximum entropy exploration,.

DIAL: Distribution-Informed Adaptive Learning of Multi-Task Constraints for Safety-Critical Systems Provably efficient maximum entropy exploration,

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T00:48:34.022731Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-10T00:48:32.150961Z digest=sha256:a811ea6850568ebb14d412efdda9b4b7d103c37f3a594c4e66f8ecc4499c30ac

Observation 46912794-0df4-4cfd-b9fc-38c387092c1b · outbound

This paper cites Efficient Exploration via State Marginal Matching.

DIAL: Distribution-Informed Adaptive Learning of Multi-Task Constraints for Safety-Critical Systems Efficient Exploration via State Marginal Matching

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-10T00:48:32.155017Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T00:48:32.155017Z digest=sha256:115a526f8c26cc6036c801fff28c368fd11cb8d6066d4ada724463c28f0d29b6

Observation 501e5005-bd15-45ff-98e7-714bcc3fe0c9 · outbound

This paper cites Constrained cross-entropy method for safe reinforcement learning,.

DIAL: Distribution-Informed Adaptive Learning of Multi-Task Constraints for Safety-Critical Systems Constrained cross-entropy method for safe reinforcement learning,

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T00:48:34.005886Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-10T00:48:32.160225Z digest=sha256:7d4d4d941eb670453a3f75340fa8eb983e94fa664d924dfc9483c45906a9e9cb

Observation 8b45eaf0-623c-4d42-b538-8b8df7a11f90 · outbound

This paper cites Learning to fly,.

DIAL: Distribution-Informed Adaptive Learning of Multi-Task Constraints for Safety-Critical Systems Learning to fly,

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T00:48:33.990268Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-10T00:48:32.165376Z digest=sha256:74be2f59cc28f27ae79609a534633f7d0ba558f79562b0f11a802a1248423b04

Observation 936e9e12-8eab-4a6e-ba98-ed91ee1054d6 · outbound

This paper cites an unresolved cited work.

DIAL: Distribution-Informed Adaptive Learning of Multi-Task Constraints for Safety-Critical Systems Unresolved cited work

Reference 34

Resolution
unresolved
raw_fallback, observed 2026-08-10T00:48:33.973898Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-10T00:48:32.169801Z digest=sha256:65c44f0ffd75b77be28e7da9fc4969543424a64515a861deb7a04719f1fb39eb

Observation 3afd8016-a712-4b1d-8e5d-b78f1278f0f6 · outbound

This paper cites Apprenticeship learning via inverse rein- forcement learning,.

DIAL: Distribution-Informed Adaptive Learning of Multi-Task Constraints for Safety-Critical Systems Apprenticeship learning via inverse rein- forcement learning,

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-10T00:48:32.173834Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T00:48:32.173834Z digest=sha256:c9e72eb7d5741a4393bb3c763952db1562b064b3f2f43892b2c61f6168578b90

Observation b225c898-29e6-4348-a94b-32eaa65f2a24 · outbound

This paper cites Generative adversarial imitation learning,.

DIAL: Distribution-Informed Adaptive Learning of Multi-Task Constraints for Safety-Critical Systems Generative adversarial imitation learning,

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-10T00:48:32.178103Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T00:48:32.178103Z digest=sha256:61b17e10eeb725c75e6098269d18355b73a733450f3dfc571a408019eadc82b9

Observation 75b8c992-6b1e-4709-a1be-9e0eeed87ed5 · outbound

This paper cites Learning robust rewards with adverserial inverse reinforcement learning,.

DIAL: Distribution-Informed Adaptive Learning of Multi-Task Constraints for Safety-Critical Systems Learning robust rewards with adverserial inverse reinforcement learning,

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T00:48:33.939014Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-10T00:48:32.182217Z digest=sha256:fe4a8ee123b64b53f5a821c24da0af3c07575ea2bb48472c6fbf9865947bb953

Observation fa645a3a-0e85-4e29-8d65-0e5fea072805 · outbound

This paper cites Iq-learn: Inverse soft-q learning for imitation,.

DIAL: Distribution-Informed Adaptive Learning of Multi-Task Constraints for Safety-Critical Systems Iq-learn: Inverse soft-q learning for imitation,

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T00:48:33.923032Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-10T00:48:32.186768Z digest=sha256:4855f8cf7d7ef17bc487802b5061e25ba3bc1f9286301ac9b8e3be564548fabd

Observation 1816aead-b213-472b-a638-1e29af456356 · outbound

This paper cites Maximum entropy inverse reinforcement learning.

DIAL: Distribution-Informed Adaptive Learning of Multi-Task Constraints for Safety-Critical Systems Maximum entropy inverse reinforcement learning

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-10T00:48:32.191737Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T00:48:32.191737Z digest=sha256:409c3e7d5319995299dc643d0a3c1b007c4841ed4d77d516a1e8d5706652e88b

Observation 96faf043-56f6-454f-a84d-9e76eaea414d · outbound

This paper cites Bridging the gap between imitation learning and inverse reinforcement learning,.

DIAL: Distribution-Informed Adaptive Learning of Multi-Task Constraints for Safety-Critical Systems Bridging the gap between imitation learning and inverse reinforcement learning,

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T00:48:33.813002Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-10T00:48:32.196399Z digest=sha256:4bd41bc5def16837b96dedd77dafed7775d6b2f8159d8e106f1c677eabccd11a

Observation f68e6c6e-c5ea-496a-8125-565e9aa65e62 · outbound

This paper cites A Connection between Generative Adversarial Networks, Inverse Reinforcement Learning, and Energy-Based Models.

DIAL: Distribution-Informed Adaptive Learning of Multi-Task Constraints for Safety-Critical Systems A Connection between Generative Adversarial Networks, Inverse Reinforcement Learning, and Energy-Based Models

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-10T00:48:32.201198Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T00:48:32.201198Z digest=sha256:8bff3e96eb1503d40c5967a2a4a48e5acff4b18570d1800e7a7a19435282b38e

Observation 80cb5e59-5b8b-49eb-86b1-83b401e1e587 · outbound

This paper cites Maximum Likelihood Constraint Inference for Inverse Reinforcement Learning.

DIAL: Distribution-Informed Adaptive Learning of Multi-Task Constraints for Safety-Critical Systems Maximum Likelihood Constraint Inference for Inverse Reinforcement Learning

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-10T00:48:32.206554Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T00:48:32.206554Z digest=sha256:6901d1e33d5dbe452553560ae950608b235966a532792cac0d540308c4e06f95

Observation fd1ded94-c38f-457b-bc72-e9148432ad45 · outbound

This paper cites Learning constraints from demon- strations with grid and parametric representations,.

DIAL: Distribution-Informed Adaptive Learning of Multi-Task Constraints for Safety-Critical Systems Learning constraints from demon- strations with grid and parametric representations,

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T00:48:33.715976Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-10T00:48:32.211728Z digest=sha256:1540865303051c08a43793993e57538fc880d9b64a583a191af2677443176add

Observation 573c1ce4-7b3d-4253-b4f9-f626d4ed3881 · outbound

This paper cites Gaussian process constraint learning for scalable chance-constrained motion planning from demon- strations,.

DIAL: Distribution-Informed Adaptive Learning of Multi-Task Constraints for Safety-Critical Systems Gaussian process constraint learning for scalable chance-constrained motion planning from demon- strations,

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T00:48:33.596229Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-10T00:48:32.216974Z digest=sha256:357cbf469fb90f40d65c46833246f5b2ec150644b9139d2dc47bd578f6d7828b

Observation 1b5e8fdf-dd69-4838-92cb-de460f28a0b2 · outbound

This paper cites Learning soft constraints from constrained expert demonstrations,.

DIAL: Distribution-Informed Adaptive Learning of Multi-Task Constraints for Safety-Critical Systems Learning soft constraints from constrained expert demonstrations,

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T00:48:33.580418Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-10T00:48:32.222108Z digest=sha256:7cf7cf40acd0aea62799f7cb1daa6f6279199e2ad02a9815c657bc8e4efcd89b

Observation 1423062b-be2c-489e-8bdd-aea74a493fd5 · outbound

This paper cites Uncertainty-aware constraint inference in inverse constrained reinforcement learning,.

DIAL: Distribution-Informed Adaptive Learning of Multi-Task Constraints for Safety-Critical Systems Uncertainty-aware constraint inference in inverse constrained reinforcement learning,

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T00:48:33.564589Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-10T00:48:32.269252Z digest=sha256:b7291fc02215f2f862276ac3b68e01450962c760abe9914fde5d3295a0a63126

Observation b2a25c70-9fb6-4c3b-8337-364794e6982c · outbound

This paper cites Confidence Aware Inverse Constrained Reinforcement Learning.

DIAL: Distribution-Informed Adaptive Learning of Multi-Task Constraints for Safety-Critical Systems Confidence Aware Inverse Constrained Reinforcement Learning

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-10T00:48:32.306249Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T00:48:32.306249Z digest=sha256:882ba3366b992be74775f5b9250ea9e7869e9d8e000308b1bcff533d24e297f8

Observation cc73d1fd-dd46-48ca-829f-9081f2c4c89f · outbound

This paper cites Inverse constraint learning and gen- eralization by transferable reward decomposition,.

DIAL: Distribution-Informed Adaptive Learning of Multi-Task Constraints for Safety-Critical Systems Inverse constraint learning and gen- eralization by transferable reward decomposition,

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T00:48:33.548668Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-10T00:48:32.342237Z digest=sha256:a1c5ace7e1a3c8ae7f35eea11f9e3b09f7d6842f8a0375c40c15bde1744528a0

Observation 27802c63-be74-4062-87c6-6c7945e75c81 · outbound

This paper cites Altman, Constrained Markov decision processes.

DIAL: Distribution-Informed Adaptive Learning of Multi-Task Constraints for Safety-Critical Systems Altman, Constrained Markov decision processes

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-10T00:48:32.455776Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T00:48:32.455776Z digest=sha256:9c091d5a57c8768bac0c7eaf0775f4efa7f8f35f9e3d2c2f8edaf93923352913

Observation db4b5a4d-393b-48bc-b55e-c00e0b9c856a · outbound

This paper cites Benchmarking Batch Deep Reinforcement Learning Algorithms.

DIAL: Distribution-Informed Adaptive Learning of Multi-Task Constraints for Safety-Critical Systems Benchmarking Batch Deep Reinforcement Learning Algorithms

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-10T00:48:32.560569Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T00:48:32.560569Z digest=sha256:c14a1a622a1b739a8c61127f6fd955d58ceaf2ad37a84305ba4a452d7aba807f

Observation 21b481c7-6aa3-4a04-ab70-bd745ffb3136 · outbound

This paper cites Approximate Robust Control of Uncertain Dynamical Systems.

DIAL: Distribution-Informed Adaptive Learning of Multi-Task Constraints for Safety-Critical Systems Approximate Robust Control of Uncertain Dynamical Systems

Reference 51

Resolution
verified exact
local_arxiv, observed 2026-08-10T00:48:32.854221Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-10T00:48:32.651094Z digest=sha256:99a4a0b70ab27601945724b531e4ce164456da180d975b3f1bd22851a182d612

Observation c3efdc1a-2d29-4ed2-9928-c5a8d085c1e7 · outbound

This paper cites Bayesian methods for constraint inference in reinforcement learning,.

DIAL: Distribution-Informed Adaptive Learning of Multi-Task Constraints for Safety-Critical Systems Bayesian methods for constraint inference in reinforcement learning,

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T00:48:33.522112Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-10T00:48:32.669783Z digest=sha256:cd12e85a16ae71f1f8368968698272099739df2be4cff8cdfd4a774b24222ff5

Observation 2a5dbb09-8942-4ec3-8318-6551f0845769 · outbound

This paper cites Risk-sensitive inverse reinforcement learning via coherent risk models.

DIAL: Distribution-Informed Adaptive Learning of Multi-Task Constraints for Safety-Critical Systems Risk-sensitive inverse reinforcement learning via coherent risk models

Reference 53

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T00:48:33.505778Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-10T00:48:32.684075Z digest=sha256:035e9d956d046dc78cb6af7577f58c8a298a6b7d5d9d539aa68ffffde6f9b256

Observation a0cbd5af-f9a4-4862-814c-e8e6f8d22f31 · outbound

This paper cites Benchmarking constraint inference in inverse reinforcement learning,.

DIAL: Distribution-Informed Adaptive Learning of Multi-Task Constraints for Safety-Critical Systems Benchmarking constraint inference in inverse reinforcement learning,

Reference 54

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T00:48:33.489286Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-10T00:48:32.690639Z digest=sha256:b4da300f4d3658a963f0cbd88141691850659ebea029ff82bfbd2f259fc3516a

Observation 7150cc6a-b9b4-4b74-9a5e-8176cd2838e0 · outbound

This paper cites Near- est neighbor estimates of entropy,.

DIAL: Distribution-Informed Adaptive Learning of Multi-Task Constraints for Safety-Critical Systems Near- est neighbor estimates of entropy,

Reference 55

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T00:48:33.469520Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-10T00:48:32.701007Z digest=sha256:1ae968ece6fe45b1bd1901280f6981bf5157410f91bceefdfbdd691ab237fc0e

Observation c899864f-9de5-4213-a0f4-ef36e355e6e6 · outbound

This paper cites Particle based probability density fusion with differential shannon entropy criterion,.

DIAL: Distribution-Informed Adaptive Learning of Multi-Task Constraints for Safety-Critical Systems Particle based probability density fusion with differential shannon entropy criterion,

Reference 56

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T00:48:33.429642Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-10T00:48:32.709316Z digest=sha256:492a8b0a91807c4660b2090e44a3f4a595738a7bce7f45796e1e6e7b8976b6a3

Observation 1da520f0-4726-42aa-bf2a-1a6e291dd466 · outbound

This paper cites Task-agnostic exploration via policy gradient of a non-parametric state entropy estimate,.

DIAL: Distribution-Informed Adaptive Learning of Multi-Task Constraints for Safety-Critical Systems Task-agnostic exploration via policy gradient of a non-parametric state entropy estimate,

Reference 57

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T00:48:33.374367Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-10T00:48:32.721815Z digest=sha256:d9dbbb376937830fb8867f42b7d6214d8a73a376c62f6077d27ceb412ed8b38c

Observation 1eaa7da7-30cd-4d05-902e-ae9a0c322804 · outbound

This paper cites Reward Constrained Policy Optimization.

DIAL: Distribution-Informed Adaptive Learning of Multi-Task Constraints for Safety-Critical Systems Reward Constrained Policy Optimization

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-10T00:48:32.728968Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T00:48:32.728968Z digest=sha256:4b4794b0667248f602f29b2386ee34076f147bd387b2921d91ac9ec79ccba0ca

Observation bcf0f65e-4bce-42be-beee-e20cd5829ec7 · outbound

This paper cites An environment for autonomous driving decision- making,.

DIAL: Distribution-Informed Adaptive Learning of Multi-Task Constraints for Safety-Critical Systems An environment for autonomous driving decision- making,

Reference 59

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T00:48:33.235606Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-10T00:48:32.734110Z digest=sha256:7a30f2a2b96305ba5fb1d651c1f7b577239b9c4b949d7a0561080674f72f37d5

Observation 1cc0b4a2-166f-4fe4-a896-a1de0e1827b2 · outbound

This paper cites OpenAI Gym.

DIAL: Distribution-Informed Adaptive Learning of Multi-Task Constraints for Safety-Critical Systems OpenAI Gym

Reference 60

Resolution
unresolved
no resolver link, observed 2026-08-10T00:48:32.739075Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T00:48:32.739075Z digest=sha256:ed35a5817759fdb530f17825ba72fe0c0065edd82ed5b5f7da7b47911b1ce3af

Observation a3c275c2-1d46-4086-94d7-f557ac3a4e64 · outbound

This paper cites Adam: A Method for Stochastic Optimization.

DIAL: Distribution-Informed Adaptive Learning of Multi-Task Constraints for Safety-Critical Systems Adam: A Method for Stochastic Optimization

Reference 61

Resolution
unresolved
no resolver link, observed 2026-08-10T00:48:32.744343Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T00:48:32.744343Z digest=sha256:ee51212aee23e37489638e90d29d4c94a411d802ebd91df928e8acbc638f71d5

Observation c6810a24-0bac-4704-8bd5-de1bad7855a6 · outbound

This paper cites Constrained differential optimization,.

DIAL: Distribution-Informed Adaptive Learning of Multi-Task Constraints for Safety-Critical Systems Constrained differential optimization,

Reference 62

Resolution
unresolved
no resolver link, observed 2026-08-10T00:48:32.749342Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T00:48:32.749342Z digest=sha256:b61a0a55969dcfa3903623f0758f0b1018509dd2772b0f6f073746ac5c8179ff

Observation a6d67648-b7bf-4c36-935d-b8118952f541 · outbound

This paper cites Controlled text gen- eration as continuous optimization with multiple constraints,.

DIAL: Distribution-Informed Adaptive Learning of Multi-Task Constraints for Safety-Critical Systems Controlled text gen- eration as continuous optimization with multiple constraints,

Reference 63

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T00:48:33.069704Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-10T00:48:32.753805Z digest=sha256:9e138ae479d976f0c254f0be539d1cad1fac1d2680dea225ff664ac95f56a8cb

Pith citing papers

No inbound Pith citation observations are available.