Pith. sign in

Paper Citation Record · LEDGER

Prompting Decision Transformers for Zero-Shot Reach-Avoid Policies

As of 9 August 2026, this Paper Citation Record lists 61 of 61 outbound references and 0 inbound Pith citation observations for arXiv:2505.19337.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.19337 v2

Coverage vector

measured 61 of 61 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T14:21:29.435813Z

measured 61 of 61 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

61 of 61 outbound references displayed

  • verified exact1
  • verified fuzzy45
  • unresolved15
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation ae52ad05-4262-4008-a82f-9c60562da714 · outbound

This paper cites A review of reward functions for reinforcement learning in the context of autonomous driving.

Prompting Decision Transformers for Zero-Shot Reach-Avoid Policies A review of reward functions for reinforcement learning in the context of autonomous driving

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:21:36.159398Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:21:24.932108Z digest=sha256:85c391ff2220fce09c31329802fb66a1e1f7e536c06bc48cec8b026d0bd6bf7d

Observation 1495beaf-23bc-4b1f-8e60-3db0b9161aa0 · outbound

This paper cites Hindsight experience replay.

Prompting Decision Transformers for Zero-Shot Reach-Avoid Policies Hindsight experience replay

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:21:36.051685Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:21:24.991707Z digest=sha256:f207f5555d7dc5574808862a154188282e44c4074a2e42e35b45a241e8d1be5d

Observation 9c5b4ea9-7f87-4975-9c69-2d817661fef0 · outbound

This paper cites Offline Goal-Conditioned Reinforcement Learning for Safety-Critical Tasks with Recovery Policy.

Prompting Decision Transformers for Zero-Shot Reach-Avoid Policies Offline Goal-Conditioned Reinforcement Learning for Safety-Critical Tasks with Recovery Policy

Reference 3

Resolution
verified exact
local_arxiv, observed 2026-08-07T14:21:29.770963Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:21:25.084900Z digest=sha256:da0d593091234821435a1b143375004473f4bf4d22fa98fb123835bbe661b12e

Observation 933a4504-06f1-4922-9b93-77336d51342f · outbound

This paper cites Actionable Models: Unsupervised Offline Reinforcement Learning of Robotic Skills.

Prompting Decision Transformers for Zero-Shot Reach-Avoid Policies Actionable Models: Unsupervised Offline Reinforcement Learning of Robotic Skills

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-07T14:21:25.121008Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:21:25.121008Z digest=sha256:19507f7914db03dd7ba15b46e84083d99a0cb1f0f1989a566b06739945237e7a

Observation d01ee715-bf00-4edb-909b-acfd60bf7f16 · outbound

This paper cites Decision Transformer: Reinforcement Learning via Sequence Modeling.

Prompting Decision Transformers for Zero-Shot Reach-Avoid Policies Decision Transformer: Reinforcement Learning via Sequence Modeling

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-07T14:21:25.163416Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:21:25.163416Z digest=sha256:a1f88a825e9a8f6da71c8d7db6e91ec0431a60f8f3301228b57537ee52d1b0c2

Observation 457babd5-5a2a-42f9-b215-a1b260f6e891 · outbound

This paper cites Gymnasium robotics, 2024.

Prompting Decision Transformers for Zero-Shot Reach-Avoid Policies Gymnasium robotics, 2024

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-07T14:21:25.216093Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:21:25.216093Z digest=sha256:97b73336067807cc89902a598ea20822ee2cc3c6d631e1437a3defeb670540e2

Observation 8bada3ed-c54a-4ff1-a252-9f16eeec7edf · outbound

This paper cites Contrastive Learning as Goal-Conditioned Reinforcement Learning.

Prompting Decision Transformers for Zero-Shot Reach-Avoid Policies Contrastive Learning as Goal-Conditioned Reinforcement Learning

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-07T14:21:25.312416Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:21:25.312416Z digest=sha256:61e544c59f62ebc9747b5a4f7103203700e1d82d2de953a1f7dd268c178e2100

Observation 5de72c4d-b6bf-4b36-a11f-df3dd372b951 · outbound

This paper cites Safe multi-agent navigation guided by goal- conditioned safe reinforcement learning, 2025.

Prompting Decision Transformers for Zero-Shot Reach-Avoid Policies Safe multi-agent navigation guided by goal- conditioned safe reinforcement learning, 2025

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:21:35.948907Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:21:25.373485Z digest=sha256:14c1788ce956e14f24e1b85de7e2b77f535e40d13cbe748b38d6eea89e8c1205

Observation 7565752c-768f-4597-a351-4e6dc68adcab · outbound

This paper cites Curriculum reinforcement learning for complex reward functions.

Prompting Decision Transformers for Zero-Shot Reach-Avoid Policies Curriculum reinforcement learning for complex reward functions

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:21:35.792698Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:21:25.413568Z digest=sha256:340cfd1620c290a3a73da674d1d85ca7bb4945f3c78db461a8be839fcf486c3d

Observation 4dff962f-8890-4939-bd0c-985c3d70ef3a · outbound

This paper cites Off-policy deep reinforcement learning without exploration.

Prompting Decision Transformers for Zero-Shot Reach-Avoid Policies Off-policy deep reinforcement learning without exploration

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-07T14:21:25.446509Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:21:25.446509Z digest=sha256:73b21b09c8359eccb40e500afe7e551a7e0991ab1e14a4413eb2da6a00fe8c73

Observation 90562ba9-61e4-41f6-8a4f-6d9098548bb4 · outbound

This paper cites Integrating domain knowledge for handling limited data in offline RL.

Prompting Decision Transformers for Zero-Shot Reach-Avoid Policies Integrating domain knowledge for handling limited data in offline RL

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:21:35.704832Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:21:25.528577Z digest=sha256:6adea96f8d5a0f265343f27ce9708970e7d7617764c4611fe469dd11246a8199

Observation a2365892-1fd6-4786-9f4b-ca7f9256155e · outbound

This paper cites Learning to Reach Goals via Iterated Supervised Learning.

Prompting Decision Transformers for Zero-Shot Reach-Avoid Policies Learning to Reach Goals via Iterated Supervised Learning

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-07T14:21:25.585866Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:21:25.585866Z digest=sha256:7ff05e4a362a6de137430fd2c0d9d9bd1a3fc5f82d5e80351583e046d6882c5c

Observation 9970028b-05b8-44f1-b722-08857b85e611 · outbound

This paper cites Bullet-safety-gym: A framework for constrained reinforcement learning.

Prompting Decision Transformers for Zero-Shot Reach-Avoid Policies Bullet-safety-gym: A framework for constrained reinforcement learning

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-07T14:21:25.636687Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:21:25.636687Z digest=sha256:36ba8f975a3439b4eec328b9291189dde61630721ec274bf5b47f5aedb8aece4

Observation 9b630f1c-0390-4132-9913-f3b28e69f4c0 · outbound

This paper cites Chemical reprogramming of human somatic cells to pluripotent stem cells.Nature, 605(7909):325–331, May 2022.

Prompting Decision Transformers for Zero-Shot Reach-Avoid Policies Chemical reprogramming of human somatic cells to pluripotent stem cells.Nature, 605(7909):325–331, May 2022

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:21:35.618483Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:21:25.685668Z digest=sha256:0c6d79d8155289772dc4a10ddaa0ae20f903ccb1ddc30c0ede2440da9360c0eb

Observation 3dbeb6d6-8591-44b4-a41a-f01b776b3d6d · outbound

This paper cites A boolean model of the cardiac gene regulatory network determining first and second heart field identity.

Prompting Decision Transformers for Zero-Shot Reach-Avoid Policies A boolean model of the cardiac gene regulatory network determining first and second heart field identity

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:21:35.510312Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:21:25.765036Z digest=sha256:43a4ef9cd62f35aa97dac5958996a599d0394e2ea456b6c5f08b482ed2e28cfc

Observation 19a83738-c370-4fee-83f2-a72b0b7c191a · outbound

This paper cites Tomlin, and Jaime F.

Prompting Decision Transformers for Zero-Shot Reach-Avoid Policies Tomlin, and Jaime F

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:21:35.439543Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:21:25.835613Z digest=sha256:950ee7404af82798c544019a5114c8f3f4caceda5b4e1646f3dcdef2b4a8a572

Observation 41d37acb-7049-474a-8f5d-645c98de102d · outbound

This paper cites Offline reinforcement learning as one big sequence modeling problem.

Prompting Decision Transformers for Zero-Shot Reach-Avoid Policies Offline reinforcement learning as one big sequence modeling problem

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:21:35.333702Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:21:25.870439Z digest=sha256:b6aaa652f95aa3b15cd98a5a611d68e40fb0c337d6f8cd67101e9484332eaf69

Observation a6ac1ca0-9b63-442a-b3b2-3336d349cf5f · outbound

This paper cites Bradley Knox, Alessandro Allievi, Holger Banzhaf, Felix Schmitt, and Peter Stone.

Prompting Decision Transformers for Zero-Shot Reach-Avoid Policies Bradley Knox, Alessandro Allievi, Holger Banzhaf, Felix Schmitt, and Peter Stone

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:21:35.300791Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:21:25.913074Z digest=sha256:89c463797d686c1455785634d6d4c941b1ee8032b6a7b4e481b141e25df418bf

Observation 070e12ce-7d13-4cf8-a63d-73078647b43c · outbound

This paper cites Bradley Knox and James MacGlashan.

Prompting Decision Transformers for Zero-Shot Reach-Avoid Policies Bradley Knox and James MacGlashan

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:21:35.242897Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:21:25.958197Z digest=sha256:5d85263001bdb05ddb14956f263d14d182d5072cbc2b100d3ad3386466984304

Observation fde16a3c-67aa-45f5-a778-05c824c8086e · outbound

This paper cites Offline reinforcement learning with implicit q-learning.

Prompting Decision Transformers for Zero-Shot Reach-Avoid Policies Offline reinforcement learning with implicit q-learning

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:21:35.061773Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:21:26.027651Z digest=sha256:f6cf51763ce1638452aa76de93acf3d425f59a644f300f56cf8bc605a784b6bd

Observation b3a8f137-4bd1-45fb-aafd-d451f660ee51 · outbound

This paper cites Stabilizing off-policy q-learning via bootstrapping error reduction.

Prompting Decision Transformers for Zero-Shot Reach-Avoid Policies Stabilizing off-policy q-learning via bootstrapping error reduction

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:21:34.984672Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:21:26.094429Z digest=sha256:393c896e388d5163f05d56492e3ebd0495b777515e613cc6d647660252dc881d

Observation 64f57ae5-f4a7-4a64-9bd3-0a61aa870116 · outbound

This paper cites Should i run offline reinforcement learning or behavioral cloning? InInternational Conference on Learning Representations, 2022.

Prompting Decision Transformers for Zero-Shot Reach-Avoid Policies Should i run offline reinforcement learning or behavioral cloning? InInternational Conference on Learning Representations, 2022

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:21:34.867017Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:21:26.141209Z digest=sha256:4b6e12b27748d84e177fe844dac9d94846a9e93e273a8ebffe82cb3f000c9a3a

Observation f615fdb4-b6e0-4481-b0ae-1381711316b9 · outbound

This paper cites Batch policy learning under constraints.

Prompting Decision Transformers for Zero-Shot Reach-Avoid Policies Batch policy learning under constraints

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:21:34.794039Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:21:26.214577Z digest=sha256:375530c1189c59eba163aabb9a5b53385f7c6e2de9cde1525c6afdcb07890082

Observation 29940060-8782-4a84-bcbe-a71e41b9ff96 · outbound

This paper cites COptiDICE: Offline constrained reinforcement learning via stationary distribution correction estimation.

Prompting Decision Transformers for Zero-Shot Reach-Avoid Policies COptiDICE: Offline constrained reinforcement learning via stationary distribution correction estimation

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:21:34.682331Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:21:26.285379Z digest=sha256:bd8e90932da3b1ce1801c86069631b9cd9555438baa5ddaa4db5b27bdd87c673

Observation f6c4a3aa-cbd2-4447-9333-4b437eb5bc4e · outbound

This paper cites Possible strategies to reduce the tumorigenic risk of reprogrammed normal and cancer cells.Int.

Prompting Decision Transformers for Zero-Shot Reach-Avoid Policies Possible strategies to reduce the tumorigenic risk of reprogrammed normal and cancer cells.Int

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:21:34.616390Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:21:26.370187Z digest=sha256:4794188752c07d9bd2ad0d39fa4d51ab2c9045a127a944585401a41cedb87af6

Observation 3d0b3000-3afe-42dd-8183-142dda00f2bd · outbound

This paper cites Datasets and benchmarks for offline safe reinforcement learning.Journal of Data-centric Machine Learning Research, 2024.

Prompting Decision Transformers for Zero-Shot Reach-Avoid Policies Datasets and benchmarks for offline safe reinforcement learning.Journal of Data-centric Machine Learning Research, 2024

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-07T14:21:26.417572Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:21:26.417572Z digest=sha256:cdea191e17763060f858579ee0a065e08aaad1b3b96b7734d9f5beb0d6cab2a9

Observation 9e2f4673-99b0-41e4-be50-de20088c4880 · outbound

This paper cites Learning latent plans from play.Conference on Robot Learning (CoRL), 2019.

Prompting Decision Transformers for Zero-Shot Reach-Avoid Policies Learning latent plans from play.Conference on Robot Learning (CoRL), 2019

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:21:34.508157Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:21:26.461167Z digest=sha256:6bd73bfcd7e25374ea56ec5b164b8099b3fc8a5a532dc37efb8ba955347d86e1

Observation 19957aec-0bc0-4bf5-a110-ffb6ca7be307 · outbound

This paper cites Offline goal-conditioned reinforcement learning via $f$-advantage regression.

Prompting Decision Transformers for Zero-Shot Reach-Avoid Policies Offline goal-conditioned reinforcement learning via $f$-advantage regression

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:21:34.400791Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:21:26.536036Z digest=sha256:dc3473763cb0435fc1642ba7564e6ec830e96ab90a89f0313475f8f366c54eb5

Observation 55ca60e0-2bd4-4da5-86fc-100ca10fd6ec · outbound

This paper cites Offline reinforcement learning with domain-unlabeled data.

Prompting Decision Transformers for Zero-Shot Reach-Avoid Policies Offline reinforcement learning with domain-unlabeled data

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:21:34.236623Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:21:26.631423Z digest=sha256:87248c25a88703990bb1c3aa2c802860fafdfdbd41e70d0f65e7c77c35f12177

Observation 1e28bae9-e1bf-4a3f-8bdb-d59383ed5638 · outbound

This paper cites Partial cellular reprogramming: A deep dive into an emerging rejuvenation technology.Aging Cell, 23(2):e14039, February 2024.

Prompting Decision Transformers for Zero-Shot Reach-Avoid Policies Partial cellular reprogramming: A deep dive into an emerging rejuvenation technology.Aging Cell, 23(2):e14039, February 2024

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:21:34.199618Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:21:26.674719Z digest=sha256:fc90f452280709349784340bf93a61450237fa1b3c59d7a35ed2d66fda6b11f0

Observation 1505f4c1-666a-4ab6-bb5e-5a148f875d0b · outbound

This paper cites Chemical reprogramming takes the fast lane.Cell Stem Cell, 30(4):335–337, April 2023.

Prompting Decision Transformers for Zero-Shot Reach-Avoid Policies Chemical reprogramming takes the fast lane.Cell Stem Cell, 30(4):335–337, April 2023

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:21:34.059635Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:21:26.718658Z digest=sha256:f67d4a0c9f374e6ea507b63c056905834462ba3cf7a47b43ef7ae224060e4f1a

Observation fc144620-7a11-4d54-b5d4-a820045f2eeb · outbound

This paper cites Ogbench: Benchmarking offline goal-conditioned rl.

Prompting Decision Transformers for Zero-Shot Reach-Avoid Policies Ogbench: Benchmarking offline goal-conditioned rl

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-07T14:21:26.828085Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:21:26.828085Z digest=sha256:32672200db3608907f7458343e2884e948d691b037462e6966e1323fd82db9a0

Observation f16168be-be60-4fb6-9d09-e87d9224aff0 · outbound

This paper cites HIQL: Offline goal- conditioned RL with latent states as actions.

Prompting Decision Transformers for Zero-Shot Reach-Avoid Policies HIQL: Offline goal- conditioned RL with latent states as actions

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:21:33.941328Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:21:26.912477Z digest=sha256:f4c586acfb1d83e5628608b67e8884ee7d8469092ec1f3bffc3e58630f7d75b4

Observation 3863d5e1-c4f1-480e-9a33-3a17062b9eb1 · outbound

This paper cites Epigenetic reprogramming as a key to reverse ageing and increase longevity.Ageing Res.

Prompting Decision Transformers for Zero-Shot Reach-Avoid Policies Epigenetic reprogramming as a key to reverse ageing and increase longevity.Ageing Res

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:21:33.772143Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:21:26.964931Z digest=sha256:fb1f313cc130ed7fdc858703a7e02160d07b369cd41ae644a5436c37a49bfe1d

Observation c81ebcd8-9b26-41a0-90d5-05a1ab410a42 · outbound

This paper cites Benchmarking Safe Exploration in Deep Reinforcement Learning.

Prompting Decision Transformers for Zero-Shot Reach-Avoid Policies Benchmarking Safe Exploration in Deep Reinforcement Learning

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-07T14:21:27.035109Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:21:27.035109Z digest=sha256:141a12fad1fda795533b96b55d095c211e19c98fd4234204df2cb383b410ad62

Observation 0c29f1ad-9d93-4042-a881-897d0b151ce5 · outbound

This paper cites Optimizing sequential gene expression modulation for cellular reprogramming - coupled boolean modeling and reinforcement learning based method.

Prompting Decision Transformers for Zero-Shot Reach-Avoid Policies Optimizing sequential gene expression modulation for cellular reprogramming - coupled boolean modeling and reinforcement learning based method

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:21:33.626843Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:21:27.153418Z digest=sha256:747dcd930b134747505e728a25c526298894f534138b4e7a25d515787eb388dc

Observation 719c873d-b636-49ae-8ca9-83c0dc06cab5 · outbound

This paper cites Solving minimum-cost reach avoid using reinforcement learning.

Prompting Decision Transformers for Zero-Shot Reach-Avoid Policies Solving minimum-cost reach avoid using reinforcement learning

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:21:33.455537Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:21:27.204231Z digest=sha256:bd2d4c4629c27f7530c27b7b4d28d0061f3cc249e57853044a7ec8ace988187c

Observation d198c106-36b2-43fc-bf20-ee396a656646 · outbound

This paper cites Responsive safety in reinforcement learning by PID lagrangian methods.

Prompting Decision Transformers for Zero-Shot Reach-Avoid Policies Responsive safety in reinforcement learning by PID lagrangian methods

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:21:33.354472Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:21:27.244936Z digest=sha256:eae49b78f72ef0b8e1c1741171da55e0dcd3e6a14e0284402d8e87bd3f169032

Observation 22f5e84f-d15b-4419-998d-7b650fcd3c16 · outbound

This paper cites Induction of pluripotent stem cells from mouse embryonic and adult fibroblast cultures by defined factors.Cell, 126(4):663–676, August 2006.

Prompting Decision Transformers for Zero-Shot Reach-Avoid Policies Induction of pluripotent stem cells from mouse embryonic and adult fibroblast cultures by defined factors.Cell, 126(4):663–676, August 2006

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:21:33.169280Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:21:27.332020Z digest=sha256:dbe1c278dbf4fb3b2f5429dd574471b28a6265a240f499e8ebdcbbf2f43e314b

Observation 81fc6768-94aa-4fd5-b5ba-dffb72dfc4dd · outbound

This paper cites Direct neuronal reprogramming: Bridging the gap between basic science and clinical application.

Prompting Decision Transformers for Zero-Shot Reach-Avoid Policies Direct neuronal reprogramming: Bridging the gap between basic science and clinical application

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:21:33.046446Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:21:27.430495Z digest=sha256:c9ff0175f2d107cd2939277660b6075fe49738a65479371c3a7f136e5d4337c3

Observation a897a0c7-abd7-45af-acaf-cb2cb7a1331d · outbound

This paper cites Strategies and mechanisms of neuronal reprogramming.Brain Res.

Prompting Decision Transformers for Zero-Shot Reach-Avoid Policies Strategies and mechanisms of neuronal reprogramming.Brain Res

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:21:32.851143Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:21:27.552585Z digest=sha256:5e38889126f29b3c2551feb6722b78335b85dfb8867be08d0df75353c81cd670

Observation f5b567f7-8873-4b9a-8df8-4648b951a71c · outbound

This paper cites Safe decision transformer with learning-based constraints.

Prompting Decision Transformers for Zero-Shot Reach-Avoid Policies Safe decision transformer with learning-based constraints

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:21:32.666599Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:21:27.649651Z digest=sha256:1bce80372b1616a09bde7cf00e8ad885bea4d1b72f9a7ffe7bf9d1f41b4926ce

Observation 33e47295-9900-4203-ac32-39d4403a5549 · outbound

This paper cites Elastic decision transformer.

Prompting Decision Transformers for Zero-Shot Reach-Avoid Policies Elastic decision transformer

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:21:32.485416Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:21:27.751195Z digest=sha256:f5d3ed92c2156565b1927438568793b134d8abbbe4f80df1b4864ab84a4f2db0

Observation d6f17874-2bc7-4a92-b511-726a498cc945 · outbound

This paper cites Prevention of tumor risk associated with the reprogramming of human pluripotent stem cells.J.

Prompting Decision Transformers for Zero-Shot Reach-Avoid Policies Prevention of tumor risk associated with the reprogramming of human pluripotent stem cells.J

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:21:32.350234Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:21:27.829982Z digest=sha256:7d94fc60434a5f41c9c8e525e33c3cd6c8784c8b487b48ab925c698d8101a5c1

Observation ec8446de-a943-44b6-974b-92a2e14c78f0 · outbound

This paper cites Constraints penalized q-learning for safe offline reinforcement learning.Proc.

Prompting Decision Transformers for Zero-Shot Reach-Avoid Policies Constraints penalized q-learning for safe offline reinforcement learning.Proc

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:21:32.172365Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:21:27.880716Z digest=sha256:6c60b5106e80a99403768acf22209aee7e2e5fbf7e3f7d756b8313d4aea734b6

Observation 645f00e5-c53b-458b-a07e-601d437ed242 · outbound

This paper cites Joshua Tenenbaum, and Chuang Gan.

Prompting Decision Transformers for Zero-Shot Reach-Avoid Policies Joshua Tenenbaum, and Chuang Gan

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:21:32.014577Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:21:27.974547Z digest=sha256:7de1e902dff125a4b578a683d08cc351529a86abe9de7aa4a34a0050feb2ea35

Observation a09187af-56c3-43c8-87dd-ddc1188ae56d · outbound

This paper cites Rethinking goal-conditioned supervised learning and its connection to offline RL.

Prompting Decision Transformers for Zero-Shot Reach-Avoid Policies Rethinking goal-conditioned supervised learning and its connection to offline RL

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-07T14:21:28.122271Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:21:28.122271Z digest=sha256:a1aed8c4ed3b5b7b7089669a4a33c3c7ba1960d30cb1a9d053399622fc9077bd

Observation aa775184-0685-471b-b5f0-9354e9cb721b · outbound

This paper cites Swapped goal-conditioned offline reinforcement learning.

Prompting Decision Transformers for Zero-Shot Reach-Avoid Policies Swapped goal-conditioned offline reinforcement learning

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:21:31.829713Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:21:28.248875Z digest=sha256:e1f9e3da0f0eb8909b437a54ab290c14fb1f9b12ce6f0b3317a18e63f7ceb0d9

Observation 364f41cf-dbec-44fa-af09-eebf4479af4b · outbound

This paper cites Pre-trained multi-goal transformers with prompt optimization for efficient online adaptation.

Prompting Decision Transformers for Zero-Shot Reach-Avoid Policies Pre-trained multi-goal transformers with prompt optimization for efficient online adaptation

Reference 49

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:21:31.676433Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:21:28.327895Z digest=sha256:35cf33395733bf6df9caa8bda6ff28ec69b1e6188f53a3b4a85139518f4531db

Observation 26d04b46-3e7c-471b-b2be-b6dac9ba78a6 · outbound

This paper cites Online Decision Transformer.

Prompting Decision Transformers for Zero-Shot Reach-Avoid Policies Online Decision Transformer

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-07T14:21:28.427899Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:21:28.427899Z digest=sha256:e1651591bf12597f01570a41effa1ec9f0f005b17cf0b9712c98e2d1c6d45b57

Observation b0b4617a-db0e-4b89-849b-5180e6f0b2ad · outbound

This paper cites attempting.

Prompting Decision Transformers for Zero-Shot Reach-Avoid Policies attempting

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:21:31.480166Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:21:28.513558Z digest=sha256:61ae08de6680e61f7cf16d093756d0e1b8d3387a5372174279b29ee11a2af996

Observation 91714414-649b-4091-84e7-e04cc85556d9 · outbound

This paper cites Constrained markov decision processes with total cost criteria: Occupation measures and primal LP.Math.

Prompting Decision Transformers for Zero-Shot Reach-Avoid Policies Constrained markov decision processes with total cost criteria: Occupation measures and primal LP.Math

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:21:31.305262Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:21:28.593570Z digest=sha256:88749e355982a74b8434d9de003a238897a63cb4a221b73226dad6ba4c7760ad

Observation c90cf0b8-f87b-4efb-a9bf-4f07173c00f7 · outbound

This paper cites An efficient algorith for determining the convex hull of a finite planar set.Inf.

Prompting Decision Transformers for Zero-Shot Reach-Avoid Policies An efficient algorith for determining the convex hull of a finite planar set.Inf

Reference 53

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:21:31.151567Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:21:28.735155Z digest=sha256:5bcb400e7c8883a756d2bebd86e1b92effad84d20eb0ab49ae652fb252b86968

Observation 3c4ae2a4-c47f-41a2-a97f-f7b787c03a27 · outbound

This paper cites Cosine annealing with warmup for pytorch.

Prompting Decision Transformers for Zero-Shot Reach-Avoid Policies Cosine annealing with warmup for pytorch

Reference 54

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:21:31.055318Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:21:28.825129Z digest=sha256:86a75d6844c61d87f15fa0d93af054df9b5979a81c29bce9cb3befa615f2328d

Observation fc3ae9f4-ed21-4f3f-a0fd-1ad405207a86 · outbound

This paper cites Tune: A Research Platform for Distributed Model Selection and Training.

Prompting Decision Transformers for Zero-Shot Reach-Avoid Policies Tune: A Research Platform for Distributed Model Selection and Training

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-07T14:21:28.915801Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:21:28.915801Z digest=sha256:f359ae1e36f58b2c0fa530517982e57b085eec3e03f3eff3e88ac1052a8d845d

Observation 49aa24ea-f7f6-4f17-a854-2475887477c4 · outbound

This paper cites A new concave hull algorithm and concaveness measure for n-dimensional datasets.Journal of Information Science and Engineering, 29:379–392, 03 2013.

Prompting Decision Transformers for Zero-Shot Reach-Avoid Policies A new concave hull algorithm and concaveness measure for n-dimensional datasets.Journal of Information Science and Engineering, 29:379–392, 03 2013

Reference 56

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:21:30.950557Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:21:28.989109Z digest=sha256:f1a4f271f3a2a3ef0a4c68b52f17fa7581813686e710a589e44253e7f3fa6dcd

Observation 27de630f-f718-4b98-bc1a-2acde3be25cf · outbound

This paper cites Language models are unsupervised multitask learners.

Prompting Decision Transformers for Zero-Shot Reach-Avoid Policies Language models are unsupervised multitask learners

Reference 57

Resolution
unresolved
no resolver link, observed 2026-08-07T14:21:29.069477Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:21:29.069477Z digest=sha256:c4022ce9fe95148922f1d66f498cd05b1e17e3f739d33b86e83b60c72bf5e63b

Observation ee2c9eec-e71c-4554-afb9-ca7904b261f8 · outbound

This paper cites Pay attention to what matters.

Prompting Decision Transformers for Zero-Shot Reach-Avoid Policies Pay attention to what matters

Reference 58

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:21:30.825656Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:21:29.159136Z digest=sha256:404b1c818e357fd7c51012c3b17263a24e5b666a1ddee410ff0ac425aec7836d

Observation d3273ad3-ad56-4f19-ada0-778141a333ca · outbound

This paper cites concave_hull.https://github.com/cubao/concave_hull.git, 2022.

Prompting Decision Transformers for Zero-Shot Reach-Avoid Policies concave_hull.https://github.com/cubao/concave_hull.git, 2022

Reference 59

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:21:30.566152Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:21:29.274220Z digest=sha256:5bfcbc762bfef5e1f0182f1cd9f4203f4d757c64644efabeeec7c428e82c65d6

Observation 41df043f-01aa-43ac-b4bb-418911854e30 · outbound

This paper cites 11th EAI International Conference, ICCASA 2022 Vinh Long, Vietnam, October 27–28, 2022 Proceedings.04 2023.

Prompting Decision Transformers for Zero-Shot Reach-Avoid Policies 11th EAI International Conference, ICCASA 2022 Vinh Long, Vietnam, October 27–28, 2022 Proceedings.04 2023

Reference 60

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:21:30.322274Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:21:29.355717Z digest=sha256:84339dd32d09149735c2eb5ec1738acdbb954172ef11c1957efdcabb35164265

Observation 320e1e58-922e-4c2d-8f19-5b5401d74c1e · outbound

This paper cites an unresolved cited work.

Prompting Decision Transformers for Zero-Shot Reach-Avoid Policies Unresolved cited work

Reference 61

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:21:30.024476Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:21:29.435813Z digest=sha256:41de5148b61c788b1d2bbaaf2062e769b687868bb8b639ad8cd2f803aad1f56f

Pith citing papers

No inbound Pith citation observations are available.