Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-14T15:31:05.521816Z
Paper Citation Record · LEDGER
As of 16 August 2026, this Paper Citation Record lists 43 of 43 outbound references and 0 inbound Pith citation observations for arXiv:1908.01022.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-14T15:31:05.521816Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-16T06:30:59.297886+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
43 of 43 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 65b0b90a-3e02-49e0-95da-dcb52ce62544 · outbound
Health-Informed Policy Gradients for Multi-Agent Reinforcement Learning Johnson, and Mykel J
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 2abd25d9-1d41-4b62-b7a1-cba858270789 · outbound
Health-Informed Policy Gradients for Multi-Agent Reinforcement Learning Unresolved cited work
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation c3f3bea2-e71a-45a9-b433-67b050cac5b1 · outbound
Health-Informed Policy Gradients for Multi-Agent Reinforcement Learning Unresolved cited work
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation b11fd6d3-dd02-4198-b51f-42c40520d513 · outbound
Health-Informed Policy Gradients for Multi-Agent Reinforcement Learning Unresolved cited work
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 9e36de8c-4caf-4d17-8afe-26f0fa00b6f9 · outbound
Health-Informed Policy Gradients for Multi-Agent Reinforcement Learning Fast and Accurate Deep Network Learning by Exponential Linear Units (ELUs)
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 390e912c-f17c-4e01-83da-bf9735861889 · outbound
Health-Informed Policy Gradients for Multi-Agent Reinforcement Learning Unresolved cited work
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 31080e64-0030-4af8-9554-dd8f96dbe302 · outbound
Health-Informed Policy Gradients for Multi-Agent Reinforcement Learning Unresolved cited work
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation cbcaafa7-d455-418d-8f95-a939dcbc5237 · outbound
Health-Informed Policy Gradients for Multi-Agent Reinforcement Learning Clipped Action Policy Gradient
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a7b9d074-431a-4ed1-9a19-397a40aca2d9 · outbound
Health-Informed Policy Gradients for Multi-Agent Reinforcement Learning Unresolved cited work
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 77bfb7dd-02fe-43fc-aa02-000d6e612255 · outbound
Health-Informed Policy Gradients for Multi-Agent Reinforcement Learning Unresolved cited work
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 95fa0923-ad83-4470-9988-dab739bc888e · outbound
Health-Informed Policy Gradients for Multi-Agent Reinforcement Learning A Survey of Learning in Multiagent Environments: Dealing with Non-Stationarity
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 32cf939a-3c82-4e49-8946-0fa8e015955c · outbound
Health-Informed Policy Gradients for Multi-Agent Reinforcement Learning Unresolved cited work
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8f3f02ff-79d9-4035-8c09-03ce77149e4e · outbound
Health-Informed Policy Gradients for Multi-Agent Reinforcement Learning Unresolved cited work
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 32931c9f-8002-4bc8-8fd4-ebd9481ddb33 · outbound
Health-Informed Policy Gradients for Multi-Agent Reinforcement Learning Adam: A Method for Stochastic Optimization
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3720ba56-e903-4da3-bbfe-a3d0ea01153d · outbound
Health-Informed Policy Gradients for Multi-Agent Reinforcement Learning Unresolved cited work
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 367d1e24-2482-48c5-a289-cfcb8691f10a · outbound
Health-Informed Policy Gradients for Multi-Agent Reinforcement Learning Unresolved cited work
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation e6754999-4130-47a4-b56f-fe4a4e9ab889 · outbound
Health-Informed Policy Gradients for Multi-Agent Reinforcement Learning Unresolved cited work
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 25514841-e649-49eb-a425-5ff16000293e · outbound
Health-Informed Policy Gradients for Multi-Agent Reinforcement Learning Unresolved cited work
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 317920fb-7582-45ea-9fc8-1f12ddfbb102 · outbound
Health-Informed Policy Gradients for Multi-Agent Reinforcement Learning Unresolved cited work
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 2515a4c3-9227-4b91-af23-521c1ba2ff2c · outbound
Health-Informed Policy Gradients for Multi-Agent Reinforcement Learning Unresolved cited work
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation fff15ea5-9d5e-40f3-be5d-e30aadbc7c29 · outbound
Health-Informed Policy Gradients for Multi-Agent Reinforcement Learning Unresolved cited work
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation a8344aac-f8b1-42e0-9db0-9a715a16d13f · outbound
Health-Informed Policy Gradients for Multi-Agent Reinforcement Learning Unresolved cited work
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 89d0028c-8ae7-42f3-9dae-0dd1ba801994 · outbound
Health-Informed Policy Gradients for Multi-Agent Reinforcement Learning Unresolved cited work
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation f9d723fa-679b-4101-ba09-bc228dbe0f23 · outbound
Health-Informed Policy Gradients for Multi-Agent Reinforcement Learning Unresolved cited work
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation d80faefe-d94e-4247-907e-f12844b09d40 · outbound
Health-Informed Policy Gradients for Multi-Agent Reinforcement Learning QMIX: Monotonic Value Function Factorisation for Deep Multi-Agent Reinforcement Learning
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cee04415-5860-4db5-a79d-4f347956314c · outbound
Health-Informed Policy Gradients for Multi-Agent Reinforcement Learning Unresolved cited work
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation c12ff084-751d-4dd2-84a8-2d52dcc6b8e2 · outbound
Health-Informed Policy Gradients for Multi-Agent Reinforcement Learning Unresolved cited work
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1c33cf1f-9f28-4881-81d1-1df3056d797e · outbound
Health-Informed Policy Gradients for Multi-Agent Reinforcement Learning Unresolved cited work
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3a596216-928c-4a0b-9e78-04139c97f911 · outbound
Health-Informed Policy Gradients for Multi-Agent Reinforcement Learning Unresolved cited work
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 56e5fd1e-a936-41b0-8b26-a7f7442ea947 · outbound
Health-Informed Policy Gradients for Multi-Agent Reinforcement Learning Unresolved cited work
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 86865a61-7c71-4f0d-b72a-c520e833a3cc · outbound
Health-Informed Policy Gradients for Multi-Agent Reinforcement Learning Unresolved cited work
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation ef9503ff-9247-442c-9534-86bfd7386253 · outbound
Health-Informed Policy Gradients for Multi-Agent Reinforcement Learning Unresolved cited work
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 01eef5a0-2d6b-4956-b49b-91a666d4146e · outbound
Health-Informed Policy Gradients for Multi-Agent Reinforcement Learning Proximal Policy Optimization Algorithms
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 94e8f63d-5a69-4f53-b82c-520e8c7c1adb · outbound
Health-Informed Policy Gradients for Multi-Agent Reinforcement Learning PettingZoo: Gym for Multi-Agent Reinforcement Learning
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1369fbe1-dc29-407e-846b-67a129431d07 · outbound
Health-Informed Policy Gradients for Multi-Agent Reinforcement Learning Revisiting Parameter Sharing in Multi-Agent Deep Reinforcement Learning
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d17a4c03-1a2a-4d9f-b905-e33d00b3d7a6 · outbound
Health-Informed Policy Gradients for Multi-Agent Reinforcement Learning Unresolved cited work
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 9ddddd13-3209-49eb-ac6d-e6f38885c044 · outbound
Health-Informed Policy Gradients for Multi-Agent Reinforcement Learning Unresolved cited work
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 41b53b4a-d3e6-475f-97e7-f9b154727802 · outbound
Health-Informed Policy Gradients for Multi-Agent Reinforcement Learning Unresolved cited work
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 2709a228-c544-464b-bf22-7d02fea8d23b · outbound
Health-Informed Policy Gradients for Multi-Agent Reinforcement Learning ∑︁ u π(u| s𝑡, θ)· 𝑛∑︁ 𝑖=1 𝑞π(s𝑡, u)∇𝜃𝑖 log𝜋𝑖(𝑎𝑖|𝜏𝑖,𝑡,𝜃𝑖) # = Eπ
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 9a3b754f-7ca8-43d6-b5d7-bae40ef791ec · outbound
Health-Informed Policy Gradients for Multi-Agent Reinforcement Learning Unresolved cited work
Reference 41
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7d9c4ee0-fbdf-4f9c-87c5-9a2b2759cfc8 · outbound
Health-Informed Policy Gradients for Multi-Agent Reinforcement Learning In International Conference on Machine Learning (ICML)
Reference 2015
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 181958b0-4945-4f8d-b794-b7216c44b7cf · outbound
Health-Informed Policy Gradients for Multi-Agent Reinforcement Learning In International Conference on Learning Representations
Reference 2016
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 010840c6-7b1b-4124-a893-4d5449536316 · outbound
Health-Informed Policy Gradients for Multi-Agent Reinforcement Learning In Advances in Neural Information Processing Systems (NIPS)
Reference 2017
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
No inbound Pith citation observations are available.