Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-06T20:25:09.694462Z
Paper Citation Record · LEDGER
As of 9 August 2026, this Paper Citation Record lists 44 of 44 outbound references and 0 inbound Pith citation observations for arXiv:2507.03372.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-06T20:25:09.694462Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
44 of 44 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 98c75044-eebe-4946-94c5-7ce03341af9d · outbound
Action Robust Reinforcement Learning via Optimal Adversary Aware Policy Optimization Reinforcement learning based recommender systems: A survey
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 9c0cddd7-2873-4880-b392-cd6d4553824d · outbound
Action Robust Reinforcement Learning via Optimal Adversary Aware Policy Optimization Safe learning in robotics: From learning-based control to safe reinforcement learning
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 6ea39787-c015-42f7-97d4-87ef7aa49182 · outbound
Action Robust Reinforcement Learning via Optimal Adversary Aware Policy Optimization Certifiable robustness to adversarial state uncertainty in deep reinforcement learning
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation caff1573-0394-4bcb-b9fa-a300889a184c · outbound
Action Robust Reinforcement Learning via Optimal Adversary Aware Policy Optimization Maximum entropy RL (provably) solves some robust RL problems
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 7d63bf7e-49bc-4d6c-b52e-e2b52d525f58 · outbound
Action Robust Reinforcement Learning via Optimal Adversary Aware Policy Optimization Online Robustness Training for Deep Reinforcement Learning
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9943ad94-02b5-4333-bd8c-c41cc21ea8b6 · outbound
Action Robust Reinforcement Learning via Optimal Adversary Aware Policy Optimization D4RL: Datasets for Deep Data-Driven Reinforcement Learning
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation dbc958f2-f026-4e4a-ba4b-5ff51183c325 · outbound
Action Robust Reinforcement Learning via Optimal Adversary Aware Policy Optimization Addressing function approximation error in actor-critic methods
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1cba9500-13f1-4507-9147-7826c7103d52 · outbound
Action Robust Reinforcement Learning via Optimal Adversary Aware Policy Optimization Unresolved cited work
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 2dcf1263-1545-4cc2-b7d7-93d9f92998be · outbound
Action Robust Reinforcement Learning via Optimal Adversary Aware Policy Optimization Adversarial Attacks on Neural Network Policies
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9a6075f1-d532-4554-b8a6-80ae8c4bdf95 · outbound
Action Robust Reinforcement Learning via Optimal Adversary Aware Policy Optimization The 37 implementation details of proximal policy optimization
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 53f44c71-7d3a-441d-b1aa-caa90f6f63e8 · outbound
Action Robust Reinforcement Learning via Optimal Adversary Aware Policy Optimization Cleanrl: High-quality single-file implementations of deep reinforcement learning algorithms
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation e464de02-4f51-4883-a397-cc3af2a83e52 · outbound
Action Robust Reinforcement Learning via Optimal Adversary Aware Policy Optimization Challenges and countermeasures for adversarial attacks on deep reinforcement learning
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a3cfaeaf-d03f-4b78-af8a-57a8459b7479 · outbound
Action Robust Reinforcement Learning via Optimal Adversary Aware Policy Optimization Learning quadrupedal locomotion over challenging terrain
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 44e43070-404d-4db8-bdd2-7d4cb313cc98 · outbound
Action Robust Reinforcement Learning via Optimal Adversary Aware Policy Optimization Spatiotem- porally constrained action space attacks on deep reinforcement learning agents
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation a10a8786-ad85-484e-b1cc-423d7da93dcb · outbound
Action Robust Reinforcement Learning via Optimal Adversary Aware Policy Optimization Efficient adversarial training without attacking: Worst-case-aware robust reinforcement learning
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 0b0ba7e6-65fc-4a8d-b3a3-8a788f77ac76 · outbound
Action Robust Reinforcement Learning via Optimal Adversary Aware Policy Optimization Provably efficient black-box action poisoning attacks against reinforcement learning
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation c7fdbf05-08ce-4686-8335-78b600678f61 · outbound
Action Robust Reinforcement Learning via Optimal Adversary Aware Policy Optimization Towards deep learning models resistant to adversarial attacks
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation c0c52c15-98af-43dc-b174-6a9692e6d100 · outbound
Action Robust Reinforcement Learning via Optimal Adversary Aware Policy Optimization Human-level control through deep reinforcement learning
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 47840ea9-ebe9-4d31-a407-a54d4a0d794d · outbound
Action Robust Reinforcement Learning via Optimal Adversary Aware Policy Optimization Robust reinforcement learning
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 87f4595a-ac59-4e40-88b9-c9c7c1bd5633 · outbound
Action Robust Reinforcement Learning via Optimal Adversary Aware Policy Optimization Assessing transferability from simulation to reality for reinforcement learning
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 0d15a0f3-b53f-4456-b67d-291feaba836a · outbound
Action Robust Reinforcement Learning via Optimal Adversary Aware Policy Optimization Robust deep reinforcement learning through adversarial loss
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation e6659db3-eda5-44b5-b5d2-b16b34babfa5 · outbound
Action Robust Reinforcement Learning via Optimal Adversary Aware Policy Optimization Characterizing attacks on deep reinforcement learning
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 3815608d-00c1-4f3d-b6af-766565bfedc9 · outbound
Action Robust Reinforcement Learning via Optimal Adversary Aware Policy Optimization Policy teaching via environment poisoning: Training-time adversarial attacks against reinforcement learning
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 94d03389-a69b-4bc9-b3ad-29fe3d273a85 · outbound
Action Robust Reinforcement Learning via Optimal Adversary Aware Policy Optimization The security of autonomous driving: Threats, defenses, and future directions
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 3bbd293c-bd7a-49d5-ba9e-9413aa00170f · outbound
Action Robust Reinforcement Learning via Optimal Adversary Aware Policy Optimization Learning to walk in minutes using massively parallel deep reinforcement learning
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3611f0eb-36b1-410b-b54d-49b501b96fe0 · outbound
Action Robust Reinforcement Learning via Optimal Adversary Aware Policy Optimization Improving robotic machining accuracy through experimental error investigation and modular compensation.The International Journal of Advanced Manufacturing Technology, 85:3–15, 2016
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation a7d3b7ca-75f7-4bcf-8e53-e080d83928f5 · outbound
Action Robust Reinforcement Learning via Optimal Adversary Aware Policy Optimization Proximal Policy Optimization Algorithms
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 44a3e993-a13a-4438-b3c0-8aaf2e46b642 · outbound
Action Robust Reinforcement Learning via Optimal Adversary Aware Policy Optimization Towards facilitating empathic conversations in online mental health support: A reinforcement learning approach
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation b967c140-f346-4417-b157-dbb834955c42 · outbound
Action Robust Reinforcement Learning via Optimal Adversary Aware Policy Optimization Certifiably robust policy learning against adversarial multi-agent communication
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 33ec714b-5e8c-447e-b95f-3630aa23a6ec · outbound
Action Robust Reinforcement Learning via Optimal Adversary Aware Policy Optimization Certifiably robust policy learning against adversarial multi-agent communication
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation ea130755-e901-474a-8744-0245ecee41df · outbound
Action Robust Reinforcement Learning via Optimal Adversary Aware Policy Optimization Who is the strongest enemy? towards optimal and efficient evasion attacks in deep RL
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation f76b6a1e-e35c-444f-9a36-5e76b5e90792 · outbound
Action Robust Reinforcement Learning via Optimal Adversary Aware Policy Optimization Reinforcement learning: An introduction
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2b867fe0-f00b-40ac-a803-ad3a36575da2 · outbound
Action Robust Reinforcement Learning via Optimal Adversary Aware Policy Optimization Action robust reinforcement learning and applications in continuous control
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 6323ea20-c761-4d13-9fbf-80fb45deebed · outbound
Action Robust Reinforcement Learning via Optimal Adversary Aware Policy Optimization Mujoco: A physics engine for model-based control
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4086eb8d-ac23-4e96-9a29-ee8438a40774 · outbound
Action Robust Reinforcement Learning via Optimal Adversary Aware Policy Optimization Policy gradient method for robust reinforcement learning
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation e99860d5-e676-412f-b7bc-4ca207a87a21 · outbound
Action Robust Reinforcement Learning via Optimal Adversary Aware Policy Optimization CROP: Certify- ing robust policies for reinforcement learning through functional smoothing
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation ca13e76a-64db-4e37-9bd1-b275932fc9a5 · outbound
Action Robust Reinforcement Learning via Optimal Adversary Aware Policy Optimization Robust deep reinforcement learning through bootstrapped opportunistic curriculum
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation b5b0e446-6c9a-4989-a042-cd08b4221c80 · outbound
Action Robust Reinforcement Learning via Optimal Adversary Aware Policy Optimization Taac: Temporally abstract actor-critic for continuous control
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 1664edd7-da4d-4e27-bbde-ce91f91b6f18 · outbound
Action Robust Reinforcement Learning via Optimal Adversary Aware Policy Optimization Gradient surgery for multi-task learning
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 92a427cb-3d4c-4772-be7c-b26a64e2dedb · outbound
Action Robust Reinforcement Learning via Optimal Adversary Aware Policy Optimization Robust reinforcement learning on state observations with learned optimal adversary
Reference 40
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation c07a37e2-7e1f-4c66-9f37-2339c35d2183 · outbound
Action Robust Reinforcement Learning via Optimal Adversary Aware Policy Optimization Robust deep reinforcement learning against adversarial perturbations on state observations
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 9d2ed019-a327-4c39-9998-5c210206e090 · outbound
Action Robust Reinforcement Learning via Optimal Adversary Aware Policy Optimization Adaptive reward-poisoning attacks against reinforcement learning
Reference 42
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation dac71140-d879-4c66-acb0-e7b5eb92cdcb · outbound
Action Robust Reinforcement Learning via Optimal Adversary Aware Policy Optimization Sim-to-real transfer in deep reinforcement learning for robotics: a survey
Reference 43
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation cde022d4-24c3-4039-9dc4-a14562d2e012 · outbound
Action Robust Reinforcement Learning via Optimal Adversary Aware Policy Optimization Cadre: A cascade deep reinforcement learning framework for vision-based autonomous urban driving
Reference 44
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
No inbound Pith citation observations are available.