Pith. sign in

Paper Citation Record · LEDGER

Robust Shielding for Safe Reinforcement Learning

As of 5 August 2026, this Paper Citation Record lists 67 of 67 outbound references and 0 inbound Pith citation observations for arXiv:2606.00270.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2606.00270 v1

Coverage vector

measured 67 of 67 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-06-28T22:21:18.863056Z

measured 67 of 67 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-05T06:32:48.257954+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

67 of 67 outbound references displayed

  • verified exact5
  • verified fuzzy0
  • unresolved62
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation ce0577be-1b1d-4d92-8c8a-1ae23b36e3a2 · outbound

This paper cites Sutton and Andrew G.

Robust Shielding for Safe Reinforcement Learning Sutton and Andrew G

Reference 1

Resolution
unresolved
no resolver link, observed 2026-06-28T22:21:18.863056Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T22:21:18.863056Z digest=sha256:ee9b44ebfc785f463fa0a8a5488aa385fd38d971708ccfdd054e4c7fee5b7a55

Observation d17ff19c-0a4a-4d91-ae44-0429fc5f8c12 · outbound

This paper cites Bagnell, and Jan Peters.

Robust Shielding for Safe Reinforcement Learning Bagnell, and Jan Peters

Reference 2

Resolution
unresolved
no resolver link, observed 2026-06-28T22:21:18.863056Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T22:21:18.863056Z digest=sha256:f27086bfc48751cc1341a582431ede2eada84ddbf458d21905627bc81931734c

Observation 4fc5b070-1932-4047-8dce-93836db5b28a · outbound

This paper cites Playing Atari with Deep Reinforcement Learning.

Robust Shielding for Safe Reinforcement Learning Playing Atari with Deep Reinforcement Learning

Reference 3

Resolution
verified exact
local_arxiv, observed 2026-06-28T22:22:43.380661Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-06-28T22:21:18.863056Z digest=sha256:9691a19aafab78121cfe967c6c33886f65188b93c789c458d0a228e5b943d4ec

Observation 1f62b9a3-9f1b-4fc8-ba76-1b42f1523bed · outbound

This paper cites Ravi Kiran, Ibrahim Sobh, Victor Talpaert, Patrick Mannion, Ahmad A.

Robust Shielding for Safe Reinforcement Learning Ravi Kiran, Ibrahim Sobh, Victor Talpaert, Patrick Mannion, Ahmad A

Reference 4

Resolution
unresolved
no resolver link, observed 2026-06-28T22:21:18.863056Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T22:21:18.863056Z digest=sha256:ed6b997ff6f02f91b73b721cbfa5e20796972686661830a1556b24b362e5df0d

Observation 13aea4b8-6ff9-4fb0-a254-2ac77af18284 · outbound

This paper cites A comprehensive survey on safe reinforcement learning.

Robust Shielding for Safe Reinforcement Learning A comprehensive survey on safe reinforcement learning

Reference 5

Resolution
unresolved
no resolver link, observed 2026-06-28T22:21:18.863056Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T22:21:18.863056Z digest=sha256:4ace44cd10b7ef0bd9ea6227c06999bc9cca79eda19741262b347c7dce162fd9

Observation b8d2f4be-a950-4273-b4a5-796e9901a4e4 · outbound

This paper cites Safe Reinforcement Learning via Shielding.

Robust Shielding for Safe Reinforcement Learning Safe Reinforcement Learning via Shielding

Reference 6

Resolution
verified exact
local_arxiv, observed 2026-06-28T22:22:43.383279Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-06-28T22:21:18.863056Z digest=sha256:2433ad28366c65901d5a0a9f36c26f0ca1485714d3738d3e10bb6bab53b9831d

Observation 27f8097c-5285-4ae7-b313-100c7694dd89 · outbound

This paper cites an unresolved cited work.

Robust Shielding for Safe Reinforcement Learning Unresolved cited work

Reference 7

Resolution
unresolved
no resolver link, observed 2026-06-28T22:21:18.863056Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T22:21:18.863056Z digest=sha256:73b28c51a3beff83b55bc56304cb73e4dc869431b1b277bf190b89a7c93e7517

Observation de4b09ec-8e73-46c9-b9ce-5ff0c582983f · outbound

This paper cites Safe reinforcement learning using probabilistic shields (invited paper).

Robust Shielding for Safe Reinforcement Learning Safe reinforcement learning using probabilistic shields (invited paper)

Reference 8

Resolution
unresolved
no resolver link, observed 2026-06-28T22:21:18.863056Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T22:21:18.863056Z digest=sha256:46fcc1641da42f101b77f53f0d678934ae4b38d1594648739328e20fec3871fd

Observation 71aa8bc3-1141-45b6-afea-67922944aaaf · outbound

This paper cites Safe reinforcement learning via shielding under partial observability.

Robust Shielding for Safe Reinforcement Learning Safe reinforcement learning via shielding under partial observability

Reference 9

Resolution
unresolved
no resolver link, observed 2026-06-28T22:21:18.863056Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T22:21:18.863056Z digest=sha256:8584c0519a73a8d29683c8a1378725116cd6b2e7c375adcc147a4e8f2efc5e4f

Observation e11f6bba-57f4-40f2-9a3c-78dd67b0dd32 · outbound

This paper cites Tsitsiklis.

Robust Shielding for Safe Reinforcement Learning Tsitsiklis

Reference 10

Resolution
unresolved
no resolver link, observed 2026-06-28T22:21:18.863056Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T22:21:18.863056Z digest=sha256:54e8811cfc5fa5593b5e59ae4a772f573da0a39a8f030f1b0236998e29d5f106

Observation 5e22b803-e770-4c18-8cd5-5e9d4d352bd4 · outbound

This paper cites Robust control of Markov decision processes with uncertain transition matrices.Oper.

Robust Shielding for Safe Reinforcement Learning Robust control of Markov decision processes with uncertain transition matrices.Oper

Reference 11

Resolution
unresolved
no resolver link, observed 2026-06-28T22:21:18.863056Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T22:21:18.863056Z digest=sha256:76bd6a7840329b921d3d2a7b0d3247ea8395d085be6fe40fbff2789d14c336c8

Observation b8c5fe6d-e060-42e0-817d-3f8c2be023c7 · outbound

This paper cites Bovy, David Parker, and Nils Jansen.

Robust Shielding for Safe Reinforcement Learning Bovy, David Parker, and Nils Jansen

Reference 12

Resolution
unresolved
no resolver link, observed 2026-06-28T22:21:18.863056Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T22:21:18.863056Z digest=sha256:fe081b50fe3ad7b72e57fcaeb0bee517b2e47b678cbc8a40bca19437336b449c

Observation ac785e77-2069-47f5-93d0-3235501d7ec5 · outbound

This paper cites Robust Markov decision processes.

Robust Shielding for Safe Reinforcement Learning Robust Markov decision processes

Reference 13

Resolution
unresolved
no resolver link, observed 2026-06-28T22:21:18.863056Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T22:21:18.863056Z digest=sha256:ee9523b73930aeeac0734e0d0db4626dfb7e310c823338f0f638c02f0073d062

Observation 477b6da7-a5ae-4b82-b04d-2e9fbf2ae387 · outbound

This paper cites an unresolved cited work.

Robust Shielding for Safe Reinforcement Learning Unresolved cited work

Reference 14

Resolution
unresolved
no resolver link, observed 2026-06-28T22:21:18.863056Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T22:21:18.863056Z digest=sha256:d6e41d893dc8bc6f27c575429ef728dc83920696208b714722580e5b1e0b7222

Observation 3b0416b9-41de-49fe-a0a2-cc335b23e419 · outbound

This paper cites Hall, Zhaocong Yuan, Siqi Zhou, Jacopo Panerati, and Angela P.

Robust Shielding for Safe Reinforcement Learning Hall, Zhaocong Yuan, Siqi Zhou, Jacopo Panerati, and Angela P

Reference 15

Resolution
unresolved
no resolver link, observed 2026-06-28T22:21:18.863056Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T22:21:18.863056Z digest=sha256:a9afbb9e86e93a60afbc9308c51674bc3233fa7cf083d5eb1c6d685106dd033c

Observation 9fd790f4-3c63-4147-998e-330cf5795d89 · outbound

This paper cites A review of safe reinforcement learning: Methods, theories, and applications.IEEE Trans.

Robust Shielding for Safe Reinforcement Learning A review of safe reinforcement learning: Methods, theories, and applications.IEEE Trans

Reference 16

Resolution
unresolved
no resolver link, observed 2026-06-28T22:21:18.863056Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T22:21:18.863056Z digest=sha256:473b5010df0f45a4d189bb521f34602810710d50b828a26016090e9756b1fe16

Observation c3a10ee1-984c-4dce-9b6e-20e99262e275 · outbound

This paper cites Shielded rein- forcement learning: A review of reactive methods for safe learning.

Robust Shielding for Safe Reinforcement Learning Shielded rein- forcement learning: A review of reactive methods for safe learning

Reference 17

Resolution
unresolved
no resolver link, observed 2026-06-28T22:21:18.863056Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T22:21:18.863056Z digest=sha256:09503256b002b14fd8f6e12d482bfc39423443a5fa31b59726fd50d0b8540636

Observation b56b7022-45d2-432d-a0f0-b9c44bc49339 · outbound

This paper cites Safe reinforcement learning via probabilistic logic shields.

Robust Shielding for Safe Reinforcement Learning Safe reinforcement learning via probabilistic logic shields

Reference 18

Resolution
unresolved
no resolver link, observed 2026-06-28T22:21:18.863056Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T22:21:18.863056Z digest=sha256:de820db78e26416b4e5e2b800c66582473ac794c95487e2b2bfbb061ec390b96

Observation 8038ff36-2d01-4c6d-86b5-c0529b88f0d0 · outbound

This paper cites Safe multi-agent reinforcement learning via shielding.

Robust Shielding for Safe Reinforcement Learning Safe multi-agent reinforcement learning via shielding

Reference 19

Resolution
unresolved
no resolver link, observed 2026-06-28T22:21:18.863056Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T22:21:18.863056Z digest=sha256:8eb5303d1c30d2002f2a9bcfc192f62ba1f1eb4b0ff3d9b62bbb5c19b5d65fb5

Observation b4a8b5c5-5553-4c53-be2c-fece36050a70 · outbound

This paper cites Online shielding for reinforcement learning.Innov.

Robust Shielding for Safe Reinforcement Learning Online shielding for reinforcement learning.Innov

Reference 20

Resolution
unresolved
no resolver link, observed 2026-06-28T22:21:18.863056Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T22:21:18.863056Z digest=sha256:68f5c4342f68e7cf94373abe468659a239f771929406d636448199036f67a84c

Observation 9fcf4fe5-1195-4219-b22c-b419a9903f41 · outbound

This paper cites Shields for safe reinforcement learning.Commun.

Robust Shielding for Safe Reinforcement Learning Shields for safe reinforcement learning.Commun

Reference 21

Resolution
unresolved
no resolver link, observed 2026-06-28T22:21:18.863056Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T22:21:18.863056Z digest=sha256:5bcb8e08091a252ad96f1b43ec67b6f78e0bfac1a0db6be1af8a556d7077126f

Observation c3943bc4-385f-452b-81e2-858ded2ad5cf · outbound

This paper cites Goodall and Francesco Belardinelli.

Robust Shielding for Safe Reinforcement Learning Goodall and Francesco Belardinelli

Reference 22

Resolution
unresolved
no resolver link, observed 2026-06-28T22:21:18.863056Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T22:21:18.863056Z digest=sha256:8d1417115d5098f6c8e32f52e7cc274bf16b4323854de114770a951425ea0ff5

Observation d10d7944-72f2-4442-9fe3-87a64e7c6012 · outbound

This paper cites Safe reinforcement learning in black-box environments via adaptive shielding.

Robust Shielding for Safe Reinforcement Learning Safe reinforcement learning in black-box environments via adaptive shielding

Reference 23

Resolution
unresolved
no resolver link, observed 2026-06-28T22:21:18.863056Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T22:21:18.863056Z digest=sha256:8db282b97dd4a2589dbc31853506a8bca22bcda5fded66cb5e8de9099ba509be

Observation 50f4470b-2379-488e-b072-72baa18e2604 · outbound

This paper cites Learning-based shielding for safe autonomy under unknown dynamics.

Robust Shielding for Safe Reinforcement Learning Learning-based shielding for safe autonomy under unknown dynamics

Reference 24

Resolution
unresolved
no resolver link, observed 2026-06-28T22:21:18.863056Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T22:21:18.863056Z digest=sha256:7b177cf1c9d007a0844bead9be81cda6c5d4a12e77e21eb77cb7c8bf76a57309

Observation e16d9690-8f62-4826-bd9b-138dc806323b · outbound

This paper cites Optimization-based robust permissive synthesis for interval MDPs, 2026.

Robust Shielding for Safe Reinforcement Learning Optimization-based robust permissive synthesis for interval MDPs, 2026

Reference 25

Resolution
unresolved
no resolver link, observed 2026-06-28T22:21:18.863056Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T22:21:18.863056Z digest=sha256:852866f31ab0b907d00cfb61916a006fb3cde82e755bb7e82599f3030b1c7631

Observation 73865658-3428-4a0a-8429-e360d9b22f13 · outbound

This paper cites Risk-constrained reinforcement learning with percentile risk criteria.J.

Robust Shielding for Safe Reinforcement Learning Risk-constrained reinforcement learning with percentile risk criteria.J

Reference 26

Resolution
unresolved
no resolver link, observed 2026-06-28T22:21:18.863056Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T22:21:18.863056Z digest=sha256:b6c309f541d5b9eeef113917a711cad1b64f3376fe09406041138d79128d57d1

Observation 56fc2446-974f-4350-b7b0-ba15e1a1c6df · outbound

This paper cites Responsive safety in reinforcement learning by PID lagrangian methods.

Robust Shielding for Safe Reinforcement Learning Responsive safety in reinforcement learning by PID lagrangian methods

Reference 27

Resolution
unresolved
no resolver link, observed 2026-06-28T22:21:18.863056Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T22:21:18.863056Z digest=sha256:15e139f8128620f83bd9d51ab15f0b9c67158a91c98f6a06838ca68f28257838

Observation e9e53553-aa10-42c1-873c-188a77b2310c · outbound

This paper cites Efficient policy optimization in robust constrained MDPs with iteration complexity guarantees.

Robust Shielding for Safe Reinforcement Learning Efficient policy optimization in robust constrained MDPs with iteration complexity guarantees

Reference 28

Resolution
unresolved
no resolver link, observed 2026-06-28T22:21:18.863056Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T22:21:18.863056Z digest=sha256:ea554e50e6e90009251fdb550debbe520b64888ad409f2e23dfd6ab16a93833c

Observation 06f6d7ed-143f-4930-aab7-b72ddb9c24b2 · outbound

This paper cites Robust Constrained-MDPs: Soft-Constrained Robust Policy Optimization under Model Uncertainty.

Robust Shielding for Safe Reinforcement Learning Robust Constrained-MDPs: Soft-Constrained Robust Policy Optimization under Model Uncertainty

Reference 29

Resolution
verified exact
arxiv_id, observed 2026-06-28T22:22:43.380251Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-06-28T22:21:18.863056Z digest=sha256:faefea760188ef574b315cd99b20ff7d0aa3b7f1b9368894cc50d7e2d9f75a5d

Observation a078195e-fedb-48d7-b424-15fde20ce79f · outbound

This paper cites Duéñez-Guzmán, and Mohammad Ghavamzadeh.

Robust Shielding for Safe Reinforcement Learning Duéñez-Guzmán, and Mohammad Ghavamzadeh

Reference 30

Resolution
unresolved
no resolver link, observed 2026-06-28T22:21:18.863056Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T22:21:18.863056Z digest=sha256:87d0ee1ccb738edfc9c995f1d5ce2d3f1f53f3f90be3fbee40aeef7e66762be4

Observation 3da81817-04be-4d89-afdc-8f630c1217a5 · outbound

This paper cites Lyapunov-based Safe Policy Optimization for Continuous Control.

Robust Shielding for Safe Reinforcement Learning Lyapunov-based Safe Policy Optimization for Continuous Control

Reference 31

Resolution
verified exact
local_arxiv, observed 2026-06-28T22:22:43.387873Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-06-28T22:21:18.863056Z digest=sha256:4aa5af29146139bcc5885820c4c9ff920f006e241ca93b5760211d9b03abea68

Observation b8dfe6ce-2db3-4cee-9b1b-ce4987235442 · outbound

This paper cites Schoellig, and Andreas Krause.

Robust Shielding for Safe Reinforcement Learning Schoellig, and Andreas Krause

Reference 32

Resolution
unresolved
no resolver link, observed 2026-06-28T22:21:18.863056Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T22:21:18.863056Z digest=sha256:7630cc6022b51595a59da9be1b27c93db2d639ce1cd94263ccf92482cffc7669

Observation 2b52fe5c-5ae0-47a7-92fd-8e222da09ea2 · outbound

This paper cites an unresolved cited work.

Robust Shielding for Safe Reinforcement Learning Unresolved cited work

Reference 33

Resolution
unresolved
no resolver link, observed 2026-06-28T22:21:18.863056Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T22:21:18.863056Z digest=sha256:c515bf480ae11e3d6199ce660866ca84eb434ecca3552edb21ebb9ec21f0e6d5

Observation a7fddf10-8d55-45c0-875b-4db38dc0068e · outbound

This paper cites When to trust your model: Model-based policy optimization.

Robust Shielding for Safe Reinforcement Learning When to trust your model: Model-based policy optimization

Reference 34

Resolution
unresolved
no resolver link, observed 2026-06-28T22:21:18.863056Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T22:21:18.863056Z digest=sha256:eba7a1278f8f09bcc93f40664c34a81362145cc596703d36f4b0e656257be876

Observation a315c0cd-f49d-40a6-9372-19fabc559c11 · outbound

This paper cites Trust the model where it trusts itself - model-based actor-critic with uncertainty-aware rollout adaption.

Robust Shielding for Safe Reinforcement Learning Trust the model where it trusts itself - model-based actor-critic with uncertainty-aware rollout adaption

Reference 35

Resolution
unresolved
no resolver link, observed 2026-06-28T22:21:18.863056Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T22:21:18.863056Z digest=sha256:dfb8a5fc93e447c5b2658d68e624b270c423619bf59df4acbba579bc2ed1d9c1

Observation 518d42ff-dcbc-43be-a1a2-f40cafc87c8f · outbound

This paper cites MIT press, 2008.

Robust Shielding for Safe Reinforcement Learning MIT press, 2008

Reference 36

Resolution
unresolved
no resolver link, observed 2026-06-28T22:21:18.863056Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T22:21:18.863056Z digest=sha256:1f6d416556c7e67b09bbd574709f9d0cbcc75f292c75edfebe9e930f41ecfcc6

Observation 09ead503-8ee9-4a5f-ba38-17910a180183 · outbound

This paper cites Bertsekas and Steven E.

Robust Shielding for Safe Reinforcement Learning Bertsekas and Steven E

Reference 37

Resolution
unresolved
no resolver link, observed 2026-06-28T22:21:18.863056Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T22:21:18.863056Z digest=sha256:b5f213067cf61f067fd0b311fdf615ebedb8cdb96d5564f0a77128f77a2eed82

Observation 4f8a9959-e5b4-48e4-b44e-bbe86e2e2cd2 · outbound

This paper cites The temporal logic of programs.

Robust Shielding for Safe Reinforcement Learning The temporal logic of programs

Reference 38

Resolution
unresolved
no resolver link, observed 2026-06-28T22:21:18.863056Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T22:21:18.863056Z digest=sha256:bf50b7e1440123d1ab73c929bd38b04dea41c6d491ba975dba0c0120ffc2bd14

Observation 8cc3b92b-df6b-4e99-81b1-f570eca27c26 · outbound

This paper cites Runtime verification - 17 years later.

Robust Shielding for Safe Reinforcement Learning Runtime verification - 17 years later

Reference 39

Resolution
unresolved
no resolver link, observed 2026-06-28T22:21:18.863056Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T22:21:18.863056Z digest=sha256:35dbe718dd62ec26dcc3ec074e7d593a1496201c06f603f624680f68badbd9b5

Observation 0343cb4f-7318-487d-ac06-35daf8b5a59a · outbound

This paper cites Deshmukh, Alexandre Donzé, Georgios Fainekos, Oded Maler, Dejan Nickovic, and Sriram Sankaranarayanan.

Robust Shielding for Safe Reinforcement Learning Deshmukh, Alexandre Donzé, Georgios Fainekos, Oded Maler, Dejan Nickovic, and Sriram Sankaranarayanan

Reference 40

Resolution
unresolved
no resolver link, observed 2026-06-28T22:21:18.863056Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T22:21:18.863056Z digest=sha256:79bf488312acffb185be57edba4bb95f3616b17cf7e5207ce93dbccfed18134e

Observation 16c97f2b-c520-4095-9c9c-77333648ee84 · outbound

This paper cites What are the odds? improving statistical model checking of Markov decision processes.

Robust Shielding for Safe Reinforcement Learning What are the odds? improving statistical model checking of Markov decision processes

Reference 41

Resolution
unresolved
no resolver link, observed 2026-06-28T22:21:18.863056Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T22:21:18.863056Z digest=sha256:355d0eafa910753f37ecdbdae94bc02cb009570c52010c6eea85b9ddb36af8cd

Observation 6ef820ae-eb83-4d79-bfdb-9f04e8941eed · outbound

This paper cites Data-driven abstraction and synthesis for stochastic systems with unknown dynamics.

Robust Shielding for Safe Reinforcement Learning Data-driven abstraction and synthesis for stochastic systems with unknown dynamics

Reference 42

Resolution
unresolved
no resolver link, observed 2026-06-28T22:21:18.863056Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T22:21:18.863056Z digest=sha256:b39c12ac7472f1e9a82a53ccd4ff458f9d18fe799e6180206c957f0afcd82f39

Observation efba3b2b-70d2-4a63-bab8-255ccc052119 · outbound

This paper cites Poonawala, Mariëlle Stoelinga, and Nils Jansen.

Robust Shielding for Safe Reinforcement Learning Poonawala, Mariëlle Stoelinga, and Nils Jansen

Reference 43

Resolution
unresolved
no resolver link, observed 2026-06-28T22:21:18.863056Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T22:21:18.863056Z digest=sha256:5679b72ef80cb900c22d6d7789dec481b8536143580679f917deee54bd41c3e9

Observation eb396f91-4c75-4ff8-a07a-2cccbcc923b9 · outbound

This paper cites Simão, David Parker, and Nils Jansen.

Robust Shielding for Safe Reinforcement Learning Simão, David Parker, and Nils Jansen

Reference 44

Resolution
unresolved
no resolver link, observed 2026-06-28T22:21:18.863056Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T22:21:18.863056Z digest=sha256:4f07c0f03edff9b975a4e35d061fcb2c4064eea9e0dca9acdbcb64e7be658a13

Observation 8c3ab334-fe77-4e8f-b0fb-033ed95b8268 · outbound

This paper cites Certifiably robust policies for uncertain parametric environments.

Robust Shielding for Safe Reinforcement Learning Certifiably robust policies for uncertain parametric environments

Reference 45

Resolution
unresolved
no resolver link, observed 2026-06-28T22:21:18.863056Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T22:21:18.863056Z digest=sha256:45abdaf4b1ef1e81225a36fa1df574d2a944e66fbcdf7eb96993a888dd248be7

Observation f62f4de0-3543-43b1-add9-b469b9b06be6 · outbound

This paper cites The use of confidence or fiducial limits illustrated in the case of the binomial.Biometrika, 26(4):404–413, 1934.

Robust Shielding for Safe Reinforcement Learning The use of confidence or fiducial limits illustrated in the case of the binomial.Biometrika, 26(4):404–413, 1934

Reference 46

Resolution
unresolved
no resolver link, observed 2026-06-28T22:21:18.863056Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T22:21:18.863056Z digest=sha256:c16c3063e050b94b4fb04e194f0181e2ed8e423984a0b6cbc7c7e140238547d0

Observation 44cf08c1-de43-416a-bea5-efa21038d718 · outbound

This paper cites Henzinger, Jan Kretínský, and Tatjana Petrov.

Robust Shielding for Safe Reinforcement Learning Henzinger, Jan Kretínský, and Tatjana Petrov

Reference 47

Resolution
unresolved
no resolver link, observed 2026-06-28T22:21:18.863056Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T22:21:18.863056Z digest=sha256:8eb82fd66b98fcd3b950699b9c3f9140c958337080764062df55db76e155dc28

Observation 92fa07e0-bf59-4b3d-9956-cf71ad6f6ba7 · outbound

This paper cites an unresolved cited work.

Robust Shielding for Safe Reinforcement Learning Unresolved cited work

Reference 48

Resolution
unresolved
no resolver link, observed 2026-06-28T22:21:18.863056Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T22:21:18.863056Z digest=sha256:1841099c58417f2d69b0a131516f9ebc6da2c4f841fea14e04cb615b93cb616d

Observation a7004950-2790-431e-aaa4-a931f3c42957 · outbound

This paper cites Proximal Policy Optimization Algorithms.

Robust Shielding for Safe Reinforcement Learning Proximal Policy Optimization Algorithms

Reference 49

Resolution
verified exact
local_arxiv, observed 2026-06-28T22:22:43.385736Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-06-28T22:21:18.863056Z digest=sha256:219c5ee41d732084ddb17a15bfc881e1467afea58d0510e079fb31f16c51daf3

Observation 013af4a1-7c88-48c3-a46d-bbbebbad59e1 · outbound

This paper cites Goodall, Omar Adalat, Edwin Hamel De-le Court, and Francesco Belardinelli.

Robust Shielding for Safe Reinforcement Learning Goodall, Omar Adalat, Edwin Hamel De-le Court, and Francesco Belardinelli

Reference 50

Resolution
unresolved
no resolver link, observed 2026-06-28T22:21:18.863056Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T22:21:18.863056Z digest=sha256:26063dfe0ad8e056b6cd4ce0340799538f77300dec16e9b10418ee25bf417be5

Observation f19afac9-e3aa-4d41-9bca-740a78fddb1d · outbound

This paper cites Simão, Marnix Suilen, and Nils Jansen.

Robust Shielding for Safe Reinforcement Learning Simão, Marnix Suilen, and Nils Jansen

Reference 51

Resolution
unresolved
no resolver link, observed 2026-06-28T22:21:18.863056Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T22:21:18.863056Z digest=sha256:eac27970ab6d85a8ac7055a36bea78eaa5fbfbb40fbc84729b23348ce1d588aa

Observation a327b1a5-80d3-4740-b03d-94987c771f81 · outbound

This paper cites Bertsimas and D.

Robust Shielding for Safe Reinforcement Learning Bertsimas and D

Reference 52

Resolution
unresolved
no resolver link, observed 2026-06-28T22:21:18.863056Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T22:21:18.863056Z digest=sha256:84c8abd9b8c8c85cf9f95afbfc299a3703400d14b6371ee6706baef19e3ea7cf

Observation 06753848-a5d4-4be4-afbb-176d5dfb7dba · outbound

This paper cites DOPE: doubly optimistic and pessimistic exploration for safe reinforcement learning.

Robust Shielding for Safe Reinforcement Learning DOPE: doubly optimistic and pessimistic exploration for safe reinforcement learning

Reference 53

Resolution
unresolved
no resolver link, observed 2026-06-28T22:21:18.863056Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T22:21:18.863056Z digest=sha256:77e065c929fde6497550df06fda5d479b21718b2dc137253e212977d03229e8a

Observation 5834b1d4-67bc-4e2e-bea9-51a76364e849 · outbound

This paper cites an unresolved cited work.

Robust Shielding for Safe Reinforcement Learning Unresolved cited work

Reference 54

Resolution
unresolved
no resolver link, observed 2026-06-28T22:21:18.863056Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T22:21:18.863056Z digest=sha256:9c55ae6bc705f69c5d840b00b4feff713f3376b05894fbc013b8c24ffca89d2c

Observation 06bc0e3f-019d-4fcb-a4c1-7c434cc3d5f5 · outbound

This paper cites an unresolved cited work.

Robust Shielding for Safe Reinforcement Learning Unresolved cited work

Reference 55

Resolution
unresolved
no resolver link, observed 2026-06-28T22:21:18.863056Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T22:21:18.863056Z digest=sha256:d6cb853660ddc8e48390cb2cd9384118bbcbccb678e5c901177c4af0980300f3

Observation 93a95186-870d-4ba1-9a42-67a3e48fa688 · outbound

This paper cites an unresolved cited work.

Robust Shielding for Safe Reinforcement Learning Unresolved cited work

Reference 56

Resolution
unresolved
no resolver link, observed 2026-06-28T22:21:18.863056Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T22:21:18.863056Z digest=sha256:d56f4ea5559a18e23456104f1828b85adb1fc37a9862b570bc0211a5f136d60c

Observation ba2ca4e9-dad7-4364-98f6-676f1cce05a7 · outbound

This paper cites ∞X t=0 γtR(st, at) # −E s0a0···∼ν.

Robust Shielding for Safe Reinforcement Learning ∞X t=0 γtR(st, at) # −E s0a0···∼ν

Reference 57

Resolution
unresolved
no resolver link, observed 2026-06-28T22:21:18.863056Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T22:21:18.863056Z digest=sha256:340f2a34035d9f39bf5d1def5c9a626f47e8b9fb7bee9206c7de59c257dc742b

Observation 7e798b27-35fe-44b6-a991-b0ca0116a769 · outbound

This paper cites ThenSis C- 2Zϵ 1−γ , γ -optimal over(M R, α)forΦ.

Robust Shielding for Safe Reinforcement Learning ThenSis C- 2Zϵ 1−γ , γ -optimal over(M R, α)forΦ

Reference 58

Resolution
unresolved
no resolver link, observed 2026-06-28T22:21:18.863056Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T22:21:18.863056Z digest=sha256:e95dad18ce499ca092a0b08894c7532e5e17f50a0fd6f5c2a693b7b148c16ffb

Observation fd056d68-8902-4788-a731-db0f71588bdc · outbound

This paper cites ThenSis C-(2Bϵ,1)-optimal over(M R, α)forΦ.

Robust Shielding for Safe Reinforcement Learning ThenSis C-(2Bϵ,1)-optimal over(M R, α)forΦ

Reference 59

Resolution
unresolved
no resolver link, observed 2026-06-28T22:21:18.863056Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T22:21:18.863056Z digest=sha256:faa37c1e7a3a1e275dac7e8da51f8cd46fbab602ccffb282844568177792582e

Observation ab086a31-470c-4fd9-8209-f92ff16f6b19 · outbound

This paper cites an unresolved cited work.

Robust Shielding for Safe Reinforcement Learning Unresolved cited work

Reference 60

Resolution
unresolved
no resolver link, observed 2026-06-28T22:21:18.863056Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T22:21:18.863056Z digest=sha256:f16fdd3fbe7f5c1d4021cfdcce3fce5c72f10725df8be22dfdba0a830e774857

Observation 037965a4-c41e-46dd-b663-acc43c44fd91 · outbound

This paper cites In particular, the shield S(MR/α,A, β ∞) is HR-(0, γ)-optimal for every γ∈(0,1] for which the corresponding return is well-defined.

Robust Shielding for Safe Reinforcement Learning In particular, the shield S(MR/α,A, β ∞) is HR-(0, γ)-optimal for every γ∈(0,1] for which the corresponding return is well-defined

Reference 61

Resolution
unresolved
no resolver link, observed 2026-06-28T22:21:18.863056Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T22:21:18.863056Z digest=sha256:8b6945e1084ee4316fd8474a66b1e578650180973cd685db416005dca0ba76ae

Observation e32c32bf-1ee2-43d0-ad12-c5e889bdb296 · outbound

This paper cites an unresolved cited work.

Robust Shielding for Safe Reinforcement Learning Unresolved cited work

Reference 62

Resolution
unresolved
no resolver link, observed 2026-06-28T22:21:18.863056Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T22:21:18.863056Z digest=sha256:bfd7b3f7c048867d8c7780127a6ef788388504717ab5025aad56b444403a170c

Observation 2d0a5dbe-f5b2-41f5-a402-5a6807cc9f85 · outbound

This paper cites ∞X t=0 c((st, qt), at) # is the probability to reachGfollowingπ, so the constraint M, π|=P ≥1−p(φ) is equivalent to Eπ.

Robust Shielding for Safe Reinforcement Learning ∞X t=0 c((st, qt), at) # is the probability to reachGfollowingπ, so the constraint M, π|=P ≥1−p(φ) is equivalent to Eπ

Reference 63

Resolution
unresolved
no resolver link, observed 2026-06-28T22:21:18.863056Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T22:21:18.863056Z digest=sha256:30e1e52e5adbc74abca63cf03be0cce18aaac9abd0b39b4d4a4ab0396f61a077

Observation c40027d8-9c0f-4380-af57-c143dd2adc40 · outbound

This paper cites an unresolved cited work.

Robust Shielding for Safe Reinforcement Learning Unresolved cited work

Reference 64

Resolution
unresolved
no resolver link, observed 2026-06-28T22:21:18.863056Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T22:21:18.863056Z digest=sha256:40680cb3da6fed236a75e9cfb159a188d10d4bc90ffbae8ea58229168be1259b

Observation 35d7b8fb-5d08-4cf8-ba51-27f479f7641d · outbound

This paper cites an unresolved cited work.

Robust Shielding for Safe Reinforcement Learning Unresolved cited work

Reference 65

Resolution
unresolved
no resolver link, observed 2026-06-28T22:21:18.863056Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T22:21:18.863056Z digest=sha256:8a823ef890a3355751c1b7b8cc437de43acdcb30e77808d8316efb1b901e7e44

Observation c76d6fad-97e9-413c-b2c3-e7d1a0d8320d · outbound

This paper cites an unresolved cited work.

Robust Shielding for Safe Reinforcement Learning Unresolved cited work

Reference 66

Resolution
unresolved
no resolver link, observed 2026-06-28T22:21:18.863056Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T22:21:18.863056Z digest=sha256:05c32ebc86b086904138c33febbbc95c7aede3b63ad6b9994489b72560123fad

Observation 0d79d067-9576-4ac9-acf8-5c99fd3db0dc · outbound

This paper cites an unresolved cited work.

Robust Shielding for Safe Reinforcement Learning Unresolved cited work

Reference 67

Resolution
unresolved
no resolver link, observed 2026-06-28T22:21:18.863056Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T22:21:18.863056Z digest=sha256:93eb834a0a51c6f3933754620daca3eebe511ab1d5d36d17b89a3f39156aebfe

Pith citing papers

No inbound Pith citation observations are available.