Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-06-28T22:21:18.863056Z
Paper Citation Record · LEDGER
As of 5 August 2026, this Paper Citation Record lists 67 of 67 outbound references and 0 inbound Pith citation observations for arXiv:2606.00270.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-06-28T22:21:18.863056Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-05T06:32:48.257954+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
67 of 67 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation ce0577be-1b1d-4d92-8c8a-1ae23b36e3a2 · outbound
Robust Shielding for Safe Reinforcement Learning Sutton and Andrew G
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d17ff19c-0a4a-4d91-ae44-0429fc5f8c12 · outbound
Robust Shielding for Safe Reinforcement Learning Bagnell, and Jan Peters
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4fc5b070-1932-4047-8dce-93836db5b28a · outbound
Robust Shielding for Safe Reinforcement Learning Playing Atari with Deep Reinforcement Learning
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 1f62b9a3-9f1b-4fc8-ba76-1b42f1523bed · outbound
Robust Shielding for Safe Reinforcement Learning Ravi Kiran, Ibrahim Sobh, Victor Talpaert, Patrick Mannion, Ahmad A
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 13aea4b8-6ff9-4fb0-a254-2ac77af18284 · outbound
Robust Shielding for Safe Reinforcement Learning A comprehensive survey on safe reinforcement learning
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b8d2f4be-a950-4273-b4a5-796e9901a4e4 · outbound
Robust Shielding for Safe Reinforcement Learning Safe Reinforcement Learning via Shielding
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 27f8097c-5285-4ae7-b313-100c7694dd89 · outbound
Robust Shielding for Safe Reinforcement Learning Unresolved cited work
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation de4b09ec-8e73-46c9-b9ce-5ff0c582983f · outbound
Robust Shielding for Safe Reinforcement Learning Safe reinforcement learning using probabilistic shields (invited paper)
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 71aa8bc3-1141-45b6-afea-67922944aaaf · outbound
Robust Shielding for Safe Reinforcement Learning Safe reinforcement learning via shielding under partial observability
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e11f6bba-57f4-40f2-9a3c-78dd67b0dd32 · outbound
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5e22b803-e770-4c18-8cd5-5e9d4d352bd4 · outbound
Robust Shielding for Safe Reinforcement Learning Robust control of Markov decision processes with uncertain transition matrices.Oper
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b8c5fe6d-e060-42e0-817d-3f8c2be023c7 · outbound
Robust Shielding for Safe Reinforcement Learning Bovy, David Parker, and Nils Jansen
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ac785e77-2069-47f5-93d0-3235501d7ec5 · outbound
Robust Shielding for Safe Reinforcement Learning Robust Markov decision processes
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 477b6da7-a5ae-4b82-b04d-2e9fbf2ae387 · outbound
Robust Shielding for Safe Reinforcement Learning Unresolved cited work
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3b0416b9-41de-49fe-a0a2-cc335b23e419 · outbound
Robust Shielding for Safe Reinforcement Learning Hall, Zhaocong Yuan, Siqi Zhou, Jacopo Panerati, and Angela P
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9fd790f4-3c63-4147-998e-330cf5795d89 · outbound
Robust Shielding for Safe Reinforcement Learning A review of safe reinforcement learning: Methods, theories, and applications.IEEE Trans
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c3a10ee1-984c-4dce-9b6e-20e99262e275 · outbound
Robust Shielding for Safe Reinforcement Learning Shielded rein- forcement learning: A review of reactive methods for safe learning
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b56b7022-45d2-432d-a0f0-b9c44bc49339 · outbound
Robust Shielding for Safe Reinforcement Learning Safe reinforcement learning via probabilistic logic shields
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8038ff36-2d01-4c6d-86b5-c0529b88f0d0 · outbound
Robust Shielding for Safe Reinforcement Learning Safe multi-agent reinforcement learning via shielding
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b4a8b5c5-5553-4c53-be2c-fece36050a70 · outbound
Robust Shielding for Safe Reinforcement Learning Online shielding for reinforcement learning.Innov
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9fcf4fe5-1195-4219-b22c-b419a9903f41 · outbound
Robust Shielding for Safe Reinforcement Learning Shields for safe reinforcement learning.Commun
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c3943bc4-385f-452b-81e2-858ded2ad5cf · outbound
Robust Shielding for Safe Reinforcement Learning Goodall and Francesco Belardinelli
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d10d7944-72f2-4442-9fe3-87a64e7c6012 · outbound
Robust Shielding for Safe Reinforcement Learning Safe reinforcement learning in black-box environments via adaptive shielding
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 50f4470b-2379-488e-b072-72baa18e2604 · outbound
Robust Shielding for Safe Reinforcement Learning Learning-based shielding for safe autonomy under unknown dynamics
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e16d9690-8f62-4826-bd9b-138dc806323b · outbound
Robust Shielding for Safe Reinforcement Learning Optimization-based robust permissive synthesis for interval MDPs, 2026
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 73865658-3428-4a0a-8429-e360d9b22f13 · outbound
Robust Shielding for Safe Reinforcement Learning Risk-constrained reinforcement learning with percentile risk criteria.J
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 56fc2446-974f-4350-b7b0-ba15e1a1c6df · outbound
Robust Shielding for Safe Reinforcement Learning Responsive safety in reinforcement learning by PID lagrangian methods
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e9e53553-aa10-42c1-873c-188a77b2310c · outbound
Robust Shielding for Safe Reinforcement Learning Efficient policy optimization in robust constrained MDPs with iteration complexity guarantees
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 06f6d7ed-143f-4930-aab7-b72ddb9c24b2 · outbound
Robust Shielding for Safe Reinforcement Learning Robust Constrained-MDPs: Soft-Constrained Robust Policy Optimization under Model Uncertainty
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation a078195e-fedb-48d7-b424-15fde20ce79f · outbound
Robust Shielding for Safe Reinforcement Learning Duéñez-Guzmán, and Mohammad Ghavamzadeh
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3da81817-04be-4d89-afdc-8f630c1217a5 · outbound
Robust Shielding for Safe Reinforcement Learning Lyapunov-based Safe Policy Optimization for Continuous Control
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation b8dfe6ce-2db3-4cee-9b1b-ce4987235442 · outbound
Robust Shielding for Safe Reinforcement Learning Schoellig, and Andreas Krause
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2b52fe5c-5ae0-47a7-92fd-8e222da09ea2 · outbound
Robust Shielding for Safe Reinforcement Learning Unresolved cited work
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a7fddf10-8d55-45c0-875b-4db38dc0068e · outbound
Robust Shielding for Safe Reinforcement Learning When to trust your model: Model-based policy optimization
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a315c0cd-f49d-40a6-9372-19fabc559c11 · outbound
Robust Shielding for Safe Reinforcement Learning Trust the model where it trusts itself - model-based actor-critic with uncertainty-aware rollout adaption
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 518d42ff-dcbc-43be-a1a2-f40cafc87c8f · outbound
Robust Shielding for Safe Reinforcement Learning MIT press, 2008
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 09ead503-8ee9-4a5f-ba38-17910a180183 · outbound
Robust Shielding for Safe Reinforcement Learning Bertsekas and Steven E
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4f8a9959-e5b4-48e4-b44e-bbe86e2e2cd2 · outbound
Robust Shielding for Safe Reinforcement Learning The temporal logic of programs
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8cc3b92b-df6b-4e99-81b1-f570eca27c26 · outbound
Robust Shielding for Safe Reinforcement Learning Runtime verification - 17 years later
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0343cb4f-7318-487d-ac06-35daf8b5a59a · outbound
Robust Shielding for Safe Reinforcement Learning Deshmukh, Alexandre Donzé, Georgios Fainekos, Oded Maler, Dejan Nickovic, and Sriram Sankaranarayanan
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 16c97f2b-c520-4095-9c9c-77333648ee84 · outbound
Robust Shielding for Safe Reinforcement Learning What are the odds? improving statistical model checking of Markov decision processes
Reference 41
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6ef820ae-eb83-4d79-bfdb-9f04e8941eed · outbound
Robust Shielding for Safe Reinforcement Learning Data-driven abstraction and synthesis for stochastic systems with unknown dynamics
Reference 42
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation efba3b2b-70d2-4a63-bab8-255ccc052119 · outbound
Robust Shielding for Safe Reinforcement Learning Poonawala, Mariëlle Stoelinga, and Nils Jansen
Reference 43
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation eb396f91-4c75-4ff8-a07a-2cccbcc923b9 · outbound
Robust Shielding for Safe Reinforcement Learning Simão, David Parker, and Nils Jansen
Reference 44
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8c3ab334-fe77-4e8f-b0fb-033ed95b8268 · outbound
Robust Shielding for Safe Reinforcement Learning Certifiably robust policies for uncertain parametric environments
Reference 45
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f62f4de0-3543-43b1-add9-b469b9b06be6 · outbound
Robust Shielding for Safe Reinforcement Learning The use of confidence or fiducial limits illustrated in the case of the binomial.Biometrika, 26(4):404–413, 1934
Reference 46
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 44cf08c1-de43-416a-bea5-efa21038d718 · outbound
Robust Shielding for Safe Reinforcement Learning Henzinger, Jan Kretínský, and Tatjana Petrov
Reference 47
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 92fa07e0-bf59-4b3d-9956-cf71ad6f6ba7 · outbound
Robust Shielding for Safe Reinforcement Learning Unresolved cited work
Reference 48
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a7004950-2790-431e-aaa4-a931f3c42957 · outbound
Robust Shielding for Safe Reinforcement Learning Proximal Policy Optimization Algorithms
Reference 49
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 013af4a1-7c88-48c3-a46d-bbbebbad59e1 · outbound
Robust Shielding for Safe Reinforcement Learning Goodall, Omar Adalat, Edwin Hamel De-le Court, and Francesco Belardinelli
Reference 50
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f19afac9-e3aa-4d41-9bca-740a78fddb1d · outbound
Robust Shielding for Safe Reinforcement Learning Simão, Marnix Suilen, and Nils Jansen
Reference 51
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a327b1a5-80d3-4740-b03d-94987c771f81 · outbound
Robust Shielding for Safe Reinforcement Learning Bertsimas and D
Reference 52
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 06753848-a5d4-4be4-afbb-176d5dfb7dba · outbound
Robust Shielding for Safe Reinforcement Learning DOPE: doubly optimistic and pessimistic exploration for safe reinforcement learning
Reference 53
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5834b1d4-67bc-4e2e-bea9-51a76364e849 · outbound
Robust Shielding for Safe Reinforcement Learning Unresolved cited work
Reference 54
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 06bc0e3f-019d-4fcb-a4c1-7c434cc3d5f5 · outbound
Robust Shielding for Safe Reinforcement Learning Unresolved cited work
Reference 55
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 93a95186-870d-4ba1-9a42-67a3e48fa688 · outbound
Robust Shielding for Safe Reinforcement Learning Unresolved cited work
Reference 56
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ba2ca4e9-dad7-4364-98f6-676f1cce05a7 · outbound
Robust Shielding for Safe Reinforcement Learning ∞X t=0 γtR(st, at) # −E s0a0···∼ν
Reference 57
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7e798b27-35fe-44b6-a991-b0ca0116a769 · outbound
Robust Shielding for Safe Reinforcement Learning ThenSis C- 2Zϵ 1−γ , γ -optimal over(M R, α)forΦ
Reference 58
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fd056d68-8902-4788-a731-db0f71588bdc · outbound
Robust Shielding for Safe Reinforcement Learning ThenSis C-(2Bϵ,1)-optimal over(M R, α)forΦ
Reference 59
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ab086a31-470c-4fd9-8209-f92ff16f6b19 · outbound
Robust Shielding for Safe Reinforcement Learning Unresolved cited work
Reference 60
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 037965a4-c41e-46dd-b663-acc43c44fd91 · outbound
Robust Shielding for Safe Reinforcement Learning In particular, the shield S(MR/α,A, β ∞) is HR-(0, γ)-optimal for every γ∈(0,1] for which the corresponding return is well-defined
Reference 61
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e32c32bf-1ee2-43d0-ad12-c5e889bdb296 · outbound
Robust Shielding for Safe Reinforcement Learning Unresolved cited work
Reference 62
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2d0a5dbe-f5b2-41f5-a402-5a6807cc9f85 · outbound
Robust Shielding for Safe Reinforcement Learning ∞X t=0 c((st, qt), at) # is the probability to reachGfollowingπ, so the constraint M, π|=P ≥1−p(φ) is equivalent to Eπ
Reference 63
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c40027d8-9c0f-4380-af57-c143dd2adc40 · outbound
Robust Shielding for Safe Reinforcement Learning Unresolved cited work
Reference 64
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 35d7b8fb-5d08-4cf8-ba51-27f479f7641d · outbound
Robust Shielding for Safe Reinforcement Learning Unresolved cited work
Reference 65
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c76d6fad-97e9-413c-b2c3-e7d1a0d8320d · outbound
Robust Shielding for Safe Reinforcement Learning Unresolved cited work
Reference 66
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0d79d067-9576-4ac9-acf8-5c99fd3db0dc · outbound
Robust Shielding for Safe Reinforcement Learning Unresolved cited work
Reference 67
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
No inbound Pith citation observations are available.