Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-06T21:20:29.473633Z
Paper Citation Record · LEDGER
As of 8 August 2026, this Paper Citation Record lists 29 of 29 outbound references and 3 inbound Pith citation observations for arXiv:2507.00485.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-06T21:20:29.473633Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-03T20:11:29.709863Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-07-03T00:07:28.391121Z
29 of 29 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 811bcd8d-b9a8-448b-8aae-9c2e93087fd6 · outbound
PNAct: Crafting Backdoor Attacks in Safe Reinforcement Learning Constrained policy optimiza- tion
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 7d838393-106d-4239-9e0d-dc73312485de · outbound
PNAct: Crafting Backdoor Attacks in Safe Reinforcement Learning Benchmarking Batch Deep Reinforcement Learning Algorithms
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0b917f32-af8c-465d-89e6-100f6087d487 · outbound
PNAct: Crafting Backdoor Attacks in Safe Reinforcement Learning Enhancing the robustness of qmix against state-adversarial attacks.Neurocomputing, 572:127191,
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 2cb0052d-2385-4b95-8162-4339f3f95b34 · outbound
PNAct: Crafting Backdoor Attacks in Safe Reinforcement Learning Robust training in multiagent deep reinforcement learning against optimal adversary
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation cb575ed0-9d50-41b5-bbc9-64257d2373f2 · outbound
PNAct: Crafting Backdoor Attacks in Safe Reinforcement Learning Backdoor attacks on safe reinforcement learning- enabled cyber–physical systems
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation a9bf24f5-be90-40d1-9626-508c8ce35f1a · outbound
PNAct: Crafting Backdoor Attacks in Safe Reinforcement Learning Trojdrl: Evaluation of back- door attacks on deep reinforcement learning
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 95836e9a-628d-41a0-afd8-5ef045ce71b3 · outbound
PNAct: Crafting Backdoor Attacks in Safe Reinforcement Learning Con- strained variational policy optimization for safe reinforce- ment learning
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 7b450561-be08-4f34-a4e9-55a1dec67e0a · outbound
PNAct: Crafting Backdoor Attacks in Safe Reinforcement Learning Towards deep learning models resistant to adversarial attacks
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation dd5e120c-00f2-4d39-b49c-0cc62563a593 · outbound
PNAct: Crafting Backdoor Attacks in Safe Reinforcement Learning Marl sim2real transfer: Merging physical reality with digital virtuality in meta- verse
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 7b95e5ce-3118-4800-b57f-0e94d1ee0e4c · outbound
PNAct: Crafting Backdoor Attacks in Safe Reinforcement Learning Responsive safety in reinforcement learn- ing by PID lagrangian methods
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 23b3a5a9-d5fa-49ff-967e-d5a4953c7e70 · outbound
PNAct: Crafting Backdoor Attacks in Safe Reinforcement Learning Mankowitz, and Shie Mannor
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 318e91ba-365c-41eb-aefa-2ba262f9fbc5 · outbound
PNAct: Crafting Backdoor Attacks in Safe Reinforcement Learning Backdoorl: Backdoor attack against competitive reinforcement learn- ing
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 6c2d81ce-7fbd-4c6a-97b1-bac205a7e6c5 · outbound
PNAct: Crafting Backdoor Attacks in Safe Reinforcement Learning Partially observable mean field multi- agent reinforcement learning based on graph attention net- work for uav swarms
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation ca4b967f-7f06-4d96-b152-23f71921dd26 · outbound
PNAct: Crafting Backdoor Attacks in Safe Reinforcement Learning First order constrained optimization in policy space
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 97f76f8e-0a90-4f66-9a50-aacb331ddb06 · outbound
PNAct: Crafting Backdoor Attacks in Safe Reinforcement Learning A robust mean-field actor-critic rein- forcement learning against adversarial perturbations on agent states
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 3ebc93a3-82df-491e-a0d5-d1d7ab5a2b9c · outbound
PNAct: Crafting Backdoor Attacks in Safe Reinforcement Learning Safety gymna- sium: A unified safe reinforcement learning benchmark
Reference 1994
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 6f5a24a6-b521-4a36-b29c-d2e8c3d88137 · outbound
PNAct: Crafting Backdoor Attacks in Safe Reinforcement Learning Constrained policy optimiza- tion via bayesian world models
Reference 1998
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation c2955608-c556-44a5-96dd-e1e86def435f · outbound
PNAct: Crafting Backdoor Attacks in Safe Reinforcement Learning Context-aware safe reinforcement learning for non- stationary environments
Reference 2005
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 0b3580d5-cdf5-45dd-8836-748dda812fe9 · outbound
PNAct: Crafting Backdoor Attacks in Safe Reinforcement Learning Unresolved cited work
Reference 2012
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation b5722f24-2869-49cc-9651-df085326ca0d · outbound
PNAct: Crafting Backdoor Attacks in Safe Reinforcement Learning Policycleanse: Backdoor detection and mitiga- tion for competitive reinforcement learning
Reference 2015
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 67b1924f-a84d-4766-b9f7-3a45409974a0 · outbound
PNAct: Crafting Backdoor Attacks in Safe Reinforcement Learning Constrained markov decision processes with total cost criteria: Lagrangian approach and dual linear program
Reference 2017
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 66f54b19-7271-4baa-84ea-a223e1c319cc · outbound
PNAct: Crafting Backdoor Attacks in Safe Reinforcement Learning Badrl: Sparse targeted backdoor attack against reinforcement learning
Reference 2018
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 2eaad8fe-ae38-46ee-bc78-4adf6c952ebe · outbound
PNAct: Crafting Backdoor Attacks in Safe Reinforcement Learning Goodfellow, Jonathon Shlens, and Christian Szegedy
Reference 2019
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 7f1d205e-0fd2-4481-bc3e-a208e4e89b08 · outbound
PNAct: Crafting Backdoor Attacks in Safe Reinforcement Learning Accelerated Primal-Dual Policy Optimization for Safe Reinforcement Learning
Reference 2020
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c45262ad-fe4e-41fd-b2cf-23cf9243c6ec · outbound
PNAct: Crafting Backdoor Attacks in Safe Reinforcement Learning Projection-Based Constrained Policy Optimization
Reference 2021
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 07ac8b65-1c42-4f9a-994f-04fea91ff6ba · outbound
PNAct: Crafting Backdoor Attacks in Safe Reinforcement Learning An online actor–critic algorithm with function approximation for constrained markov decision processes
Reference 2022
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 3a957276-6d76-4cf5-89eb-9270ba8b54bd · outbound
PNAct: Crafting Backdoor Attacks in Safe Reinforcement Learning Robust multi- agent reinforcement learning method based on adversar- ial domain randomization for real-world dual-uav co- operation
Reference 2023
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation f538e655-37da-4bcc-b5d4-43d5ddcaee51 · outbound
PNAct: Crafting Backdoor Attacks in Safe Reinforcement Learning Risk-constrained reinforcement learning with percentile risk criteria
Reference 2024
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation a0ef48ee-e03a-45d8-917b-b50a1fb35b02 · outbound
PNAct: Crafting Backdoor Attacks in Safe Reinforcement Learning Consideration of risk in re- inforcement learning
Reference 2025
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation e9f5d610-b76f-44a1-8cda-dbe946532e22 · inbound
Dataset Poisoning Attacks on Behavioral Cloning Policies PNAct: Crafting Backdoor Attacks in Safe Reinforcement Learning
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ed111060-4680-4267-9b22-ee791dc26641 · inbound
Trojan Attacks on Neural Network Controllers for Robotic Systems PNAct: Crafting Backdoor Attacks in Safe Reinforcement Learning
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a0217317-2144-4599-956a-f6d5e1852ac3 · inbound
Safe-RULE: Safe Reinforcement UnLEarning PNAct: Crafting Backdoor Attacks in Safe Reinforcement Learning
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.