Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-10T15:42:33.315453Z
Paper Citation Record · LEDGER
As of 11 August 2026, this Paper Citation Record lists 31 of 31 outbound references and 1 inbound Pith citation observation for arXiv:2501.13727.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-10T15:42:33.315453Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-11T06:34:44.6726+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-07-12T01:58:48.241542Z
A source-named dated measurement, never combined with another source.
Source: cited_works
31 of 31 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation b3c110ce-d129-4def-ab48-464b8d88cc69 · outbound
Scalable Safe Multi-Agent Reinforcement Learning for Multi-Agent System Constrained policy optimization
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 247f7d00-5d7f-4ef9-8468-e4f1f4da5963 · outbound
Scalable Safe Multi-Agent Reinforcement Learning for Multi-Agent System Learning transferable cooperative behavior in multi-agent teams
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation fecf5c0e-12be-4748-80b6-2d97106ff844 · outbound
Scalable Safe Multi-Agent Reinforcement Learning for Multi-Agent System Safe learning in robotics: From learning-based control to safe reinforcement learning
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation b51738e4-cdc8-421b-a621-f56e76a6c044 · outbound
Scalable Safe Multi-Agent Reinforcement Learning for Multi-Agent System Drqn-based 3d obstacle avoidance with a limited field of view
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 62b65941-6f06-44f2-9c4e-99f3c5b48195 · outbound
Scalable Safe Multi-Agent Reinforcement Learning for Multi-Agent System Safe rlhf: Safe reinforcement learning from human feedback
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 31728b7e-2816-4e2a-93f3-4a31ac5411cd · outbound
Scalable Safe Multi-Agent Reinforcement Learning for Multi-Agent System Safe multi-agent reinforcement learning for multi-robot control
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 197b62d4-6670-4e47-9843-8cd07326c57f · outbound
Scalable Safe Multi-Agent Reinforcement Learning for Multi-Agent System Scalable communication for multi-agent reinforcement learning via transformer-based email mechanism
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 2a51c528-284b-4c45-a2dc-a13c61f7efc5 · outbound
Scalable Safe Multi-Agent Reinforcement Learning for Multi-Agent System Long short-term memory
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 04850b18-228f-4e7a-a0e3-f2d44c3be83b · outbound
Scalable Safe Multi-Agent Reinforcement Learning for Multi-Agent System Collision avoidance and navigation for a quadrotor swarm using end-to-end deep reinforcement learning
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation c4689a2f-b261-4ccb-aff5-082cf0b1838b · outbound
Scalable Safe Multi-Agent Reinforcement Learning for Multi-Agent System Graph convolutional reinforcement learning
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 9416445e-cbed-4af1-9b8e-cf6d3f0bc66a · outbound
Scalable Safe Multi-Agent Reinforcement Learning for Multi-Agent System Semi-supervised classification with graph convolutional networks
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation deccf4ed-a799-4387-bb9b-84db804f8ad2 · outbound
Scalable Safe Multi-Agent Reinforcement Learning for Multi-Agent System Trust region policy optimisation in multi-agent reinforcement learning
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation b9e737b9-35cb-4eb4-a5d0-ef257bbd5dc6 · outbound
Scalable Safe Multi-Agent Reinforcement Learning for Multi-Agent System Cmix: Deep multi-agent reinforcement learning with peak and average constraints
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 463dc400-d9da-4470-aea6-a03698ee8d04 · outbound
Scalable Safe Multi-Agent Reinforcement Learning for Multi-Agent System Multi-agent actor-critic for mixed cooperative-competitive environments
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation c78f18e3-2e42-4dec-8301-4402589ead52 · outbound
Scalable Safe Multi-Agent Reinforcement Learning for Multi-Agent System Deep learning for safe autonomous driving: Current challenges and future directions
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 9cbd690d-4e16-42ea-8bcc-faa99229538c · outbound
Scalable Safe Multi-Agent Reinforcement Learning for Multi-Agent System Scalable multi-agent reinforcement learning through intelligent information aggregation
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 46e0a22d-d97d-4ec0-9304-e80b70c6980d · outbound
Scalable Safe Multi-Agent Reinforcement Learning for Multi-Agent System A concise introduction to decentralized POMDPs , volume 1
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 757a0885-56dc-49e2-b780-42628ca23c2b · outbound
Scalable Safe Multi-Agent Reinforcement Learning for Multi-Agent System Dealing with Non-Stationarity in Multi-Agent Deep Reinforcement Learning
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 040f8e5b-c53b-4426-ba0d-511f051134f9 · outbound
Scalable Safe Multi-Agent Reinforcement Learning for Multi-Agent System Monotonic value function factorisation for deep multi-agent reinforcement learning
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 512aa79d-cd0c-4e7e-a916-740147faf020 · outbound
Scalable Safe Multi-Agent Reinforcement Learning for Multi-Agent System Benchmarking Batch Deep Reinforcement Learning Algorithms
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b96014ec-1f09-498d-aaf6-32e0bc00b692 · outbound
Scalable Safe Multi-Agent Reinforcement Learning for Multi-Agent System High-Dimensional Continuous Control Using Generalized Advantage Estimation
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 908368f2-f73b-4d7d-b9e8-13ab0fe776b2 · outbound
Scalable Safe Multi-Agent Reinforcement Learning for Multi-Agent System Trust Region Policy Optimization
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a9bc3ede-9b43-4fd9-9957-e816f3560f63 · outbound
Scalable Safe Multi-Agent Reinforcement Learning for Multi-Agent System Masked label prediction: Unified message passing model for semi-supervised classification
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation fb0e634b-12e5-479e-862b-a4e87210f6a6 · outbound
Scalable Safe Multi-Agent Reinforcement Learning for Multi-Agent System Learning multiagent communication with backpropagation
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 760187a1-3036-47da-9699-d97f6bf33771 · outbound
Scalable Safe Multi-Agent Reinforcement Learning for Multi-Agent System Relative distributed formation and obstacle avoidance with multi-agent reinforcement learning
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 3db52748-8219-4071-aa0d-1c598580f139 · outbound
Scalable Safe Multi-Agent Reinforcement Learning for Multi-Agent System The surprising effectiveness of ppo in cooperative multi-agent games
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 4b4f57d8-16a7-4510-ba0d-676016dd04a3 · outbound
Scalable Safe Multi-Agent Reinforcement Learning for Multi-Agent System Attention-based reinforcement learning for real-time uav semantic communication
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 16e2c2fd-c0f7-4e7a-9a45-7870a63f5901 · outbound
Scalable Safe Multi-Agent Reinforcement Learning for Multi-Agent System A survey of multi-agent deep reinforcement learning with communication
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 60dd8506-e42c-4a1b-932b-02233dff5179 · outbound
Scalable Safe Multi-Agent Reinforcement Learning for Multi-Agent System Reducing Overestimation Bias in Multi-Agent Domains Using Double Centralized Critics
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2038acb6-4e53-4feb-8c36-cd29e2a5739a · outbound
Scalable Safe Multi-Agent Reinforcement Learning for Multi-Agent System Settling the variance of multi-agent policy gradients
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 26675881-98b1-4dee-9305-8bfe0b9012de · outbound
Scalable Safe Multi-Agent Reinforcement Learning for Multi-Agent System write newline
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0f220bbd-6724-4ea3-bb7c-f7b51849ea26 · inbound
High-Precision Formation Control for Heterogeneous Multi-Robot Systems via Hierarchical Hybrid Physics-Informed Deep Reinforcement Learning Scalable Safe Multi-Agent Reinforcement Learning for Multi-Agent System
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.