Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T13:30:26.255011Z
Paper Citation Record · LEDGER
As of 8 August 2026, this Paper Citation Record lists 61 of 61 outbound references and 0 inbound Pith citation observations for arXiv:2505.21852.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T13:30:26.255011Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
61 of 61 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 5caa45fb-6d24-48eb-8d23-c96a11334fb7 · outbound
A Provable Approach for End-to-End Safe Reinforcement Learning Achiam, D
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 1ea45edf-8f6a-4412-922d-31cdef7e8a52 · outbound
A Provable Approach for End-to-End Safe Reinforcement Learning Alshiekh, R
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 4ef9eeff-cff6-48f8-a334-a5a74e082f37 · outbound
A Provable Approach for End-to-End Safe Reinforcement Learning Unresolved cited work
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 7a62189b-0e02-41e8-b30e-bfed14d055d6 · outbound
A Provable Approach for End-to-End Safe Reinforcement Learning Concrete Problems in AI Safety
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fc6ff8e2-20c2-4c3b-86b9-90bfad22a305 · outbound
A Provable Approach for End-to-End Safe Reinforcement Learning Berkenkamp, M
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 74ccda11-804f-4a36-9ef6-bd3f6e011539 · outbound
A Provable Approach for End-to-End Safe Reinforcement Learning Bhatnagar and K
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 5a2a7343-b76f-4fad-ba6e-a7991a8f7ccb · outbound
A Provable Approach for End-to-End Safe Reinforcement Learning Black, M
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 32c8fa42-bdd5-451d-86ac-026819081ba9 · outbound
A Provable Approach for End-to-End Safe Reinforcement Learning Unresolved cited work
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 8d0dea74-9e4d-4cc5-a7a2-0adfaa9148bc · outbound
A Provable Approach for End-to-End Safe Reinforcement Learning Brandfonbrener, A
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 4f6df005-2369-4bd7-a90c-d7466538cc85 · outbound
A Provable Approach for End-to-End Safe Reinforcement Learning Unresolved cited work
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation c9dddd1d-31a7-4780-9147-a4281bd3b44e · outbound
A Provable Approach for End-to-End Safe Reinforcement Learning Cheng, G
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 6e2775c4-5441-4b23-a1c7-88659cd599c9 · outbound
A Provable Approach for End-to-End Safe Reinforcement Learning Unresolved cited work
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 674dfce5-6ac5-4f54-bc8a-68d633f7205e · outbound
A Provable Approach for End-to-End Safe Reinforcement Learning Da Costa, M
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b13a745c-023b-40b8-8a72-08602e0cbedc · outbound
A Provable Approach for End-to-End Safe Reinforcement Learning Emmons, B
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 1e5e9932-f7f1-489a-bb75-3eba8f490110 · outbound
A Provable Approach for End-to-End Safe Reinforcement Learning Unresolved cited work
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 092bc828-ae16-4493-8bed-b153b1723c04 · outbound
A Provable Approach for End-to-End Safe Reinforcement Learning Fujimoto, D
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation f5e404af-41b3-4494-966c-3c5cc3b0f525 · outbound
A Provable Approach for End-to-End Safe Reinforcement Learning Fulton and A
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation a7274ed4-4e90-4f4e-83fa-9173b709f53a · outbound
A Provable Approach for End-to-End Safe Reinforcement Learning Garcıa and F
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 97fcee7b-9fd3-4b4c-8a43-f7e7b001e6aa · outbound
A Provable Approach for End-to-End Safe Reinforcement Learning Gronauer
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation c32c7df1-e8cf-4c9d-b28f-55a1f4bf4c8d · outbound
A Provable Approach for End-to-End Safe Reinforcement Learning Unresolved cited work
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation adbed44e-9a9b-4514-a306-d885ac1df103 · outbound
A Provable Approach for End-to-End Safe Reinforcement Learning DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 690ac445-dd81-405f-8f18-d5424fdf5081 · outbound
A Provable Approach for End-to-End Safe Reinforcement Learning Unresolved cited work
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation dc55872e-60ae-495a-90c6-99f555156613 · outbound
A Provable Approach for End-to-End Safe Reinforcement Learning Hambly, R
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 00fbdec3-2754-4655-acb8-94c756ce42f1 · outbound
A Provable Approach for End-to-End Safe Reinforcement Learning Unresolved cited work
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 6897be8e-84ed-4d96-868a-a3899270c6c7 · outbound
A Provable Approach for End-to-End Safe Reinforcement Learning Unresolved cited work
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 8e5fce55-3c99-4ee3-a2a8-9332af96a67c · outbound
A Provable Approach for End-to-End Safe Reinforcement Learning Unresolved cited work
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation cb9f6868-d166-4fd2-9fb6-db9ed9a6026c · outbound
A Provable Approach for End-to-End Safe Reinforcement Learning Provably Safe Reinforcement Learning: Conceptual Analysis, Survey, and Benchmarking
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 02dac6d4-7dbd-4165-993c-83a9995dff40 · outbound
A Provable Approach for End-to-End Safe Reinforcement Learning Kumar, J
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 0124c1ed-df2a-44cc-8b23-59c16ca05f04 · outbound
A Provable Approach for End-to-End Safe Reinforcement Learning Reward-Conditioned Policies
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 139671e5-860b-468f-97ce-5e25eb919d11 · outbound
A Provable Approach for End-to-End Safe Reinforcement Learning Unresolved cited work
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation febdd373-fd80-472d-9a82-065076405ea4 · outbound
A Provable Approach for End-to-End Safe Reinforcement Learning Unresolved cited work
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 6ea14459-59f1-4823-b1e6-d0a9d8720cdb · outbound
A Provable Approach for End-to-End Safe Reinforcement Learning Levine, C
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation dd2a1b61-3e23-49c2-97e9-7bb23ee939f8 · outbound
A Provable Approach for End-to-End Safe Reinforcement Learning Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation dc0847ab-7a35-4193-904f-bde708311a34 · outbound
A Provable Approach for End-to-End Safe Reinforcement Learning Unresolved cited work
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 93140e8d-fce0-41ad-a807-5a00ac9cbac0 · outbound
A Provable Approach for End-to-End Safe Reinforcement Learning Unresolved cited work
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation f0751b9c-ed33-46ac-bf04-49112ec2ab7d · outbound
A Provable Approach for End-to-End Safe Reinforcement Learning Datasets and Benchmarks for Offline Safe Reinforcement Learning
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation da4a3e35-5a01-478f-b769-9c6494246e5e · outbound
A Provable Approach for End-to-End Safe Reinforcement Learning Unresolved cited work
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation dc6ab268-0dca-444c-80f9-9dea7274a561 · outbound
A Provable Approach for End-to-End Safe Reinforcement Learning Ouyang, J
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 10e93eee-0ed5-4ce2-8e51-32cbdd160ab2 · outbound
A Provable Approach for End-to-End Safe Reinforcement Learning Safe Policies for Reinforcement Learning via Primal-Dual Methods
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 4d25c2c2-a451-4962-96da-cd2e252c5317 · outbound
A Provable Approach for End-to-End Safe Reinforcement Learning Unresolved cited work
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3871230a-2563-4852-b4bb-ec764f5f4fae · outbound
A Provable Approach for End-to-End Safe Reinforcement Learning Unresolved cited work
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 4ff7817e-ca91-4258-844d-b00f8440da8f · outbound
A Provable Approach for End-to-End Safe Reinforcement Learning Satija, P
Reference 42
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 90cddf7c-eeb2-4a7d-bc5b-4a7371aa3876 · outbound
A Provable Approach for End-to-End Safe Reinforcement Learning Reinforcement Learning Upside Down: Don't Predict Rewards -- Just Map Them to Actions
Reference 43
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation add3ead9-d96b-4ea1-8ec3-d1b2e0eb72dc · outbound
A Provable Approach for End-to-End Safe Reinforcement Learning Sootla, A
Reference 44
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation d66c4c7e-dc5a-4ada-9017-3768a69efe62 · outbound
A Provable Approach for End-to-End Safe Reinforcement Learning Srinivas, A
Reference 45
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation fc5c38cc-235b-47fc-a44a-eabe153027c2 · outbound
A Provable Approach for End-to-End Safe Reinforcement Learning Training Agents using Upside-Down Reinforcement Learning
Reference 46
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 926187b4-167e-4697-8738-b0f9704de70d · outbound
A Provable Approach for End-to-End Safe Reinforcement Learning Stooke, J
Reference 47
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation d7812556-4fad-48c9-89e9-0720b8ceea70 · outbound
A Provable Approach for End-to-End Safe Reinforcement Learning Unresolved cited work
Reference 48
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 3469f3d0-0a5b-49e6-8e86-57cabe6f424d · outbound
A Provable Approach for End-to-End Safe Reinforcement Learning Unresolved cited work
Reference 49
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 47120573-a122-4411-92b6-ec129f9b34d5 · outbound
A Provable Approach for End-to-End Safe Reinforcement Learning Turchetta, F
Reference 50
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation b1411c21-861b-4134-81b4-67a501ca461e · outbound
A Provable Approach for End-to-End Safe Reinforcement Learning Unresolved cited work
Reference 51
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d32a2ba4-93e9-441a-a0b4-beb49d816e7d · outbound
A Provable Approach for End-to-End Safe Reinforcement Learning Wachi and Y
Reference 52
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation a0062dd8-7c4c-476a-aea2-d501f869a7db · outbound
A Provable Approach for End-to-End Safe Reinforcement Learning Wachi, W
Reference 53
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 327c2ee6-cc89-40f5-8722-4dfe318abc7f · outbound
A Provable Approach for End-to-End Safe Reinforcement Learning Wachi, X
Reference 54
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 3cfc6953-bd2c-471a-80eb-5e8de9f2eb1b · outbound
A Provable Approach for End-to-End Safe Reinforcement Learning Unresolved cited work
Reference 55
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 4400f6c7-0285-4bb1-9c33-50dd9d5be1a7 · outbound
A Provable Approach for End-to-End Safe Reinforcement Learning Unresolved cited work
Reference 56
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 23b55a16-b18d-49ce-a66f-b55830d03c47 · outbound
A Provable Approach for End-to-End Safe Reinforcement Learning Yang and M
Reference 57
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 701d1300-f321-4120-aba9-54be7671f6b7 · outbound
A Provable Approach for End-to-End Safe Reinforcement Learning Unresolved cited work
Reference 58
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 92c4b7fb-3b1e-4c95-9a80-e8f6a8c5dfe6 · outbound
A Provable Approach for End-to-End Safe Reinforcement Learning Unresolved cited work
Reference 59
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation d7a6d77d-4f82-41b8-9618-8e2501cd0dcc · outbound
A Provable Approach for End-to-End Safe Reinforcement Learning Unresolved cited work
Reference 60
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 6362a78c-7d61-49b9-95da-5e8cb5b3a191 · outbound
A Provable Approach for End-to-End Safe Reinforcement Learning Let β : X → ∆(A) be a behavior policy and D := {Ξ(i)}n i=1 ∼ (Pβ)n be a collection of n i.i.d
Reference 61
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
No inbound Pith citation observations are available.