Pith. sign in

Paper Citation Record · LEDGER

Health-Informed Policy Gradients for Multi-Agent Reinforcement Learning

As of 16 August 2026, this Paper Citation Record lists 43 of 43 outbound references and 0 inbound Pith citation observations for arXiv:1908.01022.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
1908.01022 v4

Coverage vector

measured 43 of 43 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-14T15:31:05.521816Z

measured 43 of 43 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-16T06:30:59.297886+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

43 of 43 outbound references displayed

  • verified exact1
  • verified fuzzy4
  • unresolved38
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 65b0b90a-3e02-49e0-95da-dcb52ce62544 · outbound

This paper cites Johnson, and Mykel J.

Health-Informed Policy Gradients for Multi-Agent Reinforcement Learning Johnson, and Mykel J

Reference 1

Resolution
verified exact
doi, observed 2026-08-14T15:31:05.566753Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-14T15:31:05.310320Z digest=sha256:b64bd63ee94b8be687c11c405a40a4789b4ab93b8a8a12085a095d464d2e0485

Observation 2abd25d9-1d41-4b62-b7a1-cba858270789 · outbound

This paper cites an unresolved cited work.

Health-Informed Policy Gradients for Multi-Agent Reinforcement Learning Unresolved cited work

Reference 2

Resolution
unresolved
raw_fallback, observed 2026-08-14T15:31:06.175089Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-14T15:31:05.316614Z digest=sha256:baaccb7142a4f0f7d8f0c11eb91223e58c38e6e0aae114acb0914c8c7e4b9250

Observation c3f3bea2-e71a-45a9-b433-67b050cac5b1 · outbound

This paper cites an unresolved cited work.

Health-Informed Policy Gradients for Multi-Agent Reinforcement Learning Unresolved cited work

Reference 3

Resolution
unresolved
raw_fallback, observed 2026-08-14T15:31:06.162235Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-14T15:31:05.323044Z digest=sha256:cc60814cca9a45f2e9460969b649d587898d98f7f5369286e1f6e83f3eb09781

Observation b11fd6d3-dd02-4198-b51f-42c40520d513 · outbound

This paper cites an unresolved cited work.

Health-Informed Policy Gradients for Multi-Agent Reinforcement Learning Unresolved cited work

Reference 4

Resolution
unresolved
raw_fallback, observed 2026-08-14T15:31:06.148539Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-14T15:31:05.328023Z digest=sha256:9c6c848b88e953c2fc49be640daeed8b5a361cae4d12f392f0e250741d25af88

Observation 9e36de8c-4caf-4d17-8afe-26f0fa00b6f9 · outbound

This paper cites Fast and Accurate Deep Network Learning by Exponential Linear Units (ELUs).

Health-Informed Policy Gradients for Multi-Agent Reinforcement Learning Fast and Accurate Deep Network Learning by Exponential Linear Units (ELUs)

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-14T15:31:05.334928Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T15:31:05.334928Z digest=sha256:e107d11598e1bd74d9590bc9191c0c495e073f6545b583ddbb2ee5be7b2ab034

Observation 390e912c-f17c-4e01-83da-bf9735861889 · outbound

This paper cites an unresolved cited work.

Health-Informed Policy Gradients for Multi-Agent Reinforcement Learning Unresolved cited work

Reference 6

Resolution
unresolved
raw_fallback, observed 2026-08-14T15:31:06.134754Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-14T15:31:05.340214Z digest=sha256:491bfbe6447fd45ead215238a8b3034ee778b71e30abb3a6060d5fac8ed77fa7

Observation 31080e64-0030-4af8-9554-dd8f96dbe302 · outbound

This paper cites an unresolved cited work.

Health-Informed Policy Gradients for Multi-Agent Reinforcement Learning Unresolved cited work

Reference 7

Resolution
unresolved
raw_fallback, observed 2026-08-14T15:31:06.120944Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-14T15:31:05.346065Z digest=sha256:aa626ec5ae5d8779adcfb1d057b8bcaec7be5a44848cfb0cc6800f087c75bcbc

Observation cbcaafa7-d455-418d-8f95-a939dcbc5237 · outbound

This paper cites Clipped Action Policy Gradient.

Health-Informed Policy Gradients for Multi-Agent Reinforcement Learning Clipped Action Policy Gradient

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-14T15:31:05.351418Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T15:31:05.351418Z digest=sha256:893edfa7dd7da881fdbae558501dd219ac6d660d98efeae5284294ea870fe240

Observation a7b9d074-431a-4ed1-9a19-397a40aca2d9 · outbound

This paper cites an unresolved cited work.

Health-Informed Policy Gradients for Multi-Agent Reinforcement Learning Unresolved cited work

Reference 9

Resolution
unresolved
raw_fallback, observed 2026-08-14T15:31:06.107449Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-14T15:31:05.358713Z digest=sha256:b0a8a3f17b28c4b2b5eb0b11968b4d49ff61ea6ebf90460eecbd2d2712641142

Observation 77bfb7dd-02fe-43fc-aa02-000d6e612255 · outbound

This paper cites an unresolved cited work.

Health-Informed Policy Gradients for Multi-Agent Reinforcement Learning Unresolved cited work

Reference 10

Resolution
unresolved
raw_fallback, observed 2026-08-14T15:31:06.093176Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-14T15:31:05.364252Z digest=sha256:9dff7cb153e79e924d6a39f94aebcecb9a5f8eca9eadfdd894c6d7c76bde34dd

Observation 95fa0923-ad83-4470-9988-dab739bc888e · outbound

This paper cites A Survey of Learning in Multiagent Environments: Dealing with Non-Stationarity.

Health-Informed Policy Gradients for Multi-Agent Reinforcement Learning A Survey of Learning in Multiagent Environments: Dealing with Non-Stationarity

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-14T15:31:05.371369Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T15:31:05.371369Z digest=sha256:dab62404b318763a896a4974eaaf8be6064c4a776548b661ad3734df27680781

Observation 32cf939a-3c82-4e49-8946-0fa8e015955c · outbound

This paper cites an unresolved cited work.

Health-Informed Policy Gradients for Multi-Agent Reinforcement Learning Unresolved cited work

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-14T15:31:05.375880Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T15:31:05.375880Z digest=sha256:af2d869a33a69b13a096bdc160cde08e2f9f7159c163faa3ee6bf7c3b8b19a29

Observation 8f3f02ff-79d9-4035-8c09-03ce77149e4e · outbound

This paper cites an unresolved cited work.

Health-Informed Policy Gradients for Multi-Agent Reinforcement Learning Unresolved cited work

Reference 13

Resolution
unresolved
raw_fallback, observed 2026-08-14T15:31:06.067606Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-14T15:31:05.380130Z digest=sha256:319f6404f569de0a097e672e3f6359c140637f8bc46282f5a60853b652eba10c

Observation 32931c9f-8002-4bc8-8fd4-ebd9481ddb33 · outbound

This paper cites Adam: A Method for Stochastic Optimization.

Health-Informed Policy Gradients for Multi-Agent Reinforcement Learning Adam: A Method for Stochastic Optimization

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-14T15:31:05.384134Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T15:31:05.384134Z digest=sha256:c4eb351fe31dd78a0765f97d8f66b24090805ad390f18e575c90c9987bdc0422

Observation 3720ba56-e903-4da3-bbfe-a3d0ea01153d · outbound

This paper cites an unresolved cited work.

Health-Informed Policy Gradients for Multi-Agent Reinforcement Learning Unresolved cited work

Reference 15

Resolution
unresolved
raw_fallback, observed 2026-08-14T15:31:06.050520Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-14T15:31:05.388898Z digest=sha256:83d0d4e2ac7de77901fd9e0d6a5e7a19c00845b0f00da15663ac29e67bdb9111

Observation 367d1e24-2482-48c5-a289-cfcb8691f10a · outbound

This paper cites an unresolved cited work.

Health-Informed Policy Gradients for Multi-Agent Reinforcement Learning Unresolved cited work

Reference 16

Resolution
unresolved
raw_fallback, observed 2026-08-14T15:31:06.035364Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-14T15:31:05.392726Z digest=sha256:fede5f58a563f2a374c918e242be360b18fc7bb3879222b52b14506aaaac06fe

Observation e6754999-4130-47a4-b56f-fe4a4e9ab889 · outbound

This paper cites an unresolved cited work.

Health-Informed Policy Gradients for Multi-Agent Reinforcement Learning Unresolved cited work

Reference 17

Resolution
unresolved
raw_fallback, observed 2026-08-14T15:31:06.022486Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-14T15:31:05.396605Z digest=sha256:0ba8675f71ad0c8aa7247af893aec7f8d21d76e340321fbc2db69789f8fcef19

Observation 25514841-e649-49eb-a425-5ff16000293e · outbound

This paper cites an unresolved cited work.

Health-Informed Policy Gradients for Multi-Agent Reinforcement Learning Unresolved cited work

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-14T15:31:05.402129Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T15:31:05.402129Z digest=sha256:e05407a03c6f2550db44ba97970cf158a47669422e5d8d70080190e5d23f8e67

Observation 317920fb-7582-45ea-9fc8-1f12ddfbb102 · outbound

This paper cites an unresolved cited work.

Health-Informed Policy Gradients for Multi-Agent Reinforcement Learning Unresolved cited work

Reference 19

Resolution
unresolved
raw_fallback, observed 2026-08-14T15:31:05.987246Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-14T15:31:05.411482Z digest=sha256:0207aca373b143a0ef57811c2a77d6d63f028681ca06b10a66a1da3cad76f8a9

Observation 2515a4c3-9227-4b91-af23-521c1ba2ff2c · outbound

This paper cites an unresolved cited work.

Health-Informed Policy Gradients for Multi-Agent Reinforcement Learning Unresolved cited work

Reference 20

Resolution
unresolved
raw_fallback, observed 2026-08-14T15:31:05.973456Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-14T15:31:05.415148Z digest=sha256:a71d51aa5a9736a9a5ee916c308f10932a340da2eb5257a6f0ba83d6883ac379

Observation fff15ea5-9d5e-40f3-be5d-e30aadbc7c29 · outbound

This paper cites an unresolved cited work.

Health-Informed Policy Gradients for Multi-Agent Reinforcement Learning Unresolved cited work

Reference 21

Resolution
unresolved
raw_fallback, observed 2026-08-14T15:31:05.957429Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-14T15:31:05.418822Z digest=sha256:843b6d979e89f7035f70c694a595e78f247bd03b956542c4aa961ab6874ba729

Observation a8344aac-f8b1-42e0-9db0-9a715a16d13f · outbound

This paper cites an unresolved cited work.

Health-Informed Policy Gradients for Multi-Agent Reinforcement Learning Unresolved cited work

Reference 22

Resolution
unresolved
raw_fallback, observed 2026-08-14T15:31:05.944517Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-14T15:31:05.422232Z digest=sha256:196b181771c8a5c6de24f7f408b4c4d19c825b85c6040df668638e4d20c4d1c2

Observation 89d0028c-8ae7-42f3-9dae-0dd1ba801994 · outbound

This paper cites an unresolved cited work.

Health-Informed Policy Gradients for Multi-Agent Reinforcement Learning Unresolved cited work

Reference 23

Resolution
unresolved
raw_fallback, observed 2026-08-14T15:31:05.929995Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-14T15:31:05.426347Z digest=sha256:95a00e945661d6eaf4d30f10f97a746efff415a09616c4926dfb6cbd5e2316ae

Observation f9d723fa-679b-4101-ba09-bc228dbe0f23 · outbound

This paper cites an unresolved cited work.

Health-Informed Policy Gradients for Multi-Agent Reinforcement Learning Unresolved cited work

Reference 24

Resolution
unresolved
raw_fallback, observed 2026-08-14T15:31:05.916570Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-14T15:31:05.429813Z digest=sha256:47e53977b7b4a91263699b2f60c53bd20255645f20c17089556ab99abf9c4fcf

Observation d80faefe-d94e-4247-907e-f12844b09d40 · outbound

This paper cites QMIX: Monotonic Value Function Factorisation for Deep Multi-Agent Reinforcement Learning.

Health-Informed Policy Gradients for Multi-Agent Reinforcement Learning QMIX: Monotonic Value Function Factorisation for Deep Multi-Agent Reinforcement Learning

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-14T15:31:05.435396Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T15:31:05.435396Z digest=sha256:7aadd042485a62fb09f4f5b98a3aa358d50a8b94e16ec427fb8be40b08a0ba68

Observation cee04415-5860-4db5-a79d-4f347956314c · outbound

This paper cites an unresolved cited work.

Health-Informed Policy Gradients for Multi-Agent Reinforcement Learning Unresolved cited work

Reference 26

Resolution
unresolved
raw_fallback, observed 2026-08-14T15:31:05.901447Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-14T15:31:05.439753Z digest=sha256:2d11f3dd54df90caf53bd50b82a435500f64c74434d9ae11e4b7b8456e8a363f

Observation c12ff084-751d-4dd2-84a8-2d52dcc6b8e2 · outbound

This paper cites an unresolved cited work.

Health-Informed Policy Gradients for Multi-Agent Reinforcement Learning Unresolved cited work

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-14T15:31:05.443913Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T15:31:05.443913Z digest=sha256:7b8f52a9d23c4034de03afb5ddf0cbd0502b5f59352556c0228bf4e15a6d4df8

Observation 1c33cf1f-9f28-4881-81d1-1df3056d797e · outbound

This paper cites an unresolved cited work.

Health-Informed Policy Gradients for Multi-Agent Reinforcement Learning Unresolved cited work

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-14T15:31:05.452852Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T15:31:05.452852Z digest=sha256:6067a80dc07f3657d55bad213c7e8dc325c63e80375b8a78219d54a0f66b250f

Observation 3a596216-928c-4a0b-9e78-04139c97f911 · outbound

This paper cites an unresolved cited work.

Health-Informed Policy Gradients for Multi-Agent Reinforcement Learning Unresolved cited work

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-14T15:31:05.461962Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T15:31:05.461962Z digest=sha256:755d0793d7f3492bb930159cf63c9938a8738bfd3bc88ed702cf3ca8db73fb37

Observation 56e5fd1e-a936-41b0-8b26-a7f7442ea947 · outbound

This paper cites an unresolved cited work.

Health-Informed Policy Gradients for Multi-Agent Reinforcement Learning Unresolved cited work

Reference 30

Resolution
unresolved
raw_fallback, observed 2026-08-14T15:31:05.818430Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-14T15:31:05.471601Z digest=sha256:ca5fdbcbf724c4ff355a80b7b302001c7e7a98c08a3e1b1e381f34970cc6bde7

Observation 86865a61-7c71-4f0d-b72a-c520e833a3cc · outbound

This paper cites an unresolved cited work.

Health-Informed Policy Gradients for Multi-Agent Reinforcement Learning Unresolved cited work

Reference 31

Resolution
unresolved
raw_fallback, observed 2026-08-14T15:31:05.805674Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-14T15:31:05.476887Z digest=sha256:bbee587388bae37e8d84307affb4af8cc225d9267ccc2a0e100712d3f72bf380

Observation ef9503ff-9247-442c-9534-86bfd7386253 · outbound

This paper cites an unresolved cited work.

Health-Informed Policy Gradients for Multi-Agent Reinforcement Learning Unresolved cited work

Reference 32

Resolution
unresolved
raw_fallback, observed 2026-08-14T15:31:05.790913Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-14T15:31:05.481402Z digest=sha256:aa4e2e8db23a02021239c3b66eb6cac75a6b35ff52eb2eb1a83fb827b10e405c

Observation 01eef5a0-2d6b-4956-b49b-91a666d4146e · outbound

This paper cites Proximal Policy Optimization Algorithms.

Health-Informed Policy Gradients for Multi-Agent Reinforcement Learning Proximal Policy Optimization Algorithms

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-14T15:31:05.466656Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T15:31:05.466656Z digest=sha256:f757468636dfbd956cc65b4bfac100aedfc64cc11c2dcf0e26063b9aa235f0a2

Observation 94e8f63d-5a69-4f53-b82c-520e8c7c1adb · outbound

This paper cites PettingZoo: Gym for Multi-Agent Reinforcement Learning.

Health-Informed Policy Gradients for Multi-Agent Reinforcement Learning PettingZoo: Gym for Multi-Agent Reinforcement Learning

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-14T15:31:05.493756Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T15:31:05.493756Z digest=sha256:4fb0b0c391cd1ddcabec5b17e7432cacb72794aeae50326d0420e5c40ca8e4bb

Observation 1369fbe1-dc29-407e-846b-67a129431d07 · outbound

This paper cites Revisiting Parameter Sharing in Multi-Agent Deep Reinforcement Learning.

Health-Informed Policy Gradients for Multi-Agent Reinforcement Learning Revisiting Parameter Sharing in Multi-Agent Deep Reinforcement Learning

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-14T15:31:05.499450Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T15:31:05.499450Z digest=sha256:dd9871f6009b31304b8f8524df849e98c9271dc4e592b7f2387af57b49efc7dd

Observation d17a4c03-1a2a-4d9f-b905-e33d00b3d7a6 · outbound

This paper cites an unresolved cited work.

Health-Informed Policy Gradients for Multi-Agent Reinforcement Learning Unresolved cited work

Reference 36

Resolution
unresolved
raw_fallback, observed 2026-08-14T15:31:05.754955Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-14T15:31:05.506160Z digest=sha256:08bd2d6040435a4d1677ac51719c03514a02c8b7c2569f024fe0038bfeb21947

Observation 9ddddd13-3209-49eb-ac6d-e6f38885c044 · outbound

This paper cites an unresolved cited work.

Health-Informed Policy Gradients for Multi-Agent Reinforcement Learning Unresolved cited work

Reference 37

Resolution
unresolved
raw_fallback, observed 2026-08-14T15:31:05.772770Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-14T15:31:05.487612Z digest=sha256:16619fbc29f6e41684bcce760f6ae7501c1c02664d32912940aee967d6035ffd

Observation 41b53b4a-d3e6-475f-97e7-f9b154727802 · outbound

This paper cites an unresolved cited work.

Health-Informed Policy Gradients for Multi-Agent Reinforcement Learning Unresolved cited work

Reference 38

Resolution
unresolved
raw_fallback, observed 2026-08-14T15:31:05.728186Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-14T15:31:05.516684Z digest=sha256:038475af78b85e8d34cf8679e01db249807f75196250a069515288094e0f2d63

Observation 2709a228-c544-464b-bf22-7d02fea8d23b · outbound

This paper cites ∑︁ u π(u| s𝑡, θ)· 𝑛∑︁ 𝑖=1 𝑞π(s𝑡, u)∇𝜃𝑖 log𝜋𝑖(𝑎𝑖|𝜏𝑖,𝑡,𝜃𝑖) # = Eπ.

Health-Informed Policy Gradients for Multi-Agent Reinforcement Learning ∑︁ u π(u| s𝑡, θ)· 𝑛∑︁ 𝑖=1 𝑞π(s𝑡, u)∇𝜃𝑖 log𝜋𝑖(𝑎𝑖|𝜏𝑖,𝑡,𝜃𝑖) # = Eπ

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T15:31:05.710236Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-14T15:31:05.521816Z digest=sha256:cd6760f084edd25ffe64ef47d290f730d6165efa81d6952e7954ac09fe8ba21a

Observation 9a3b754f-7ca8-43d6-b5d7-bae40ef791ec · outbound

This paper cites an unresolved cited work.

Health-Informed Policy Gradients for Multi-Agent Reinforcement Learning Unresolved cited work

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-14T15:31:05.510597Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T15:31:05.510597Z digest=sha256:ef41042f3ef25ae39226317697af40476ab8f7b461ed82ce6ee933f2efdcd66f

Observation 7d9c4ee0-fbdf-4f9c-87c5-9a2b2759cfc8 · outbound

This paper cites In International Conference on Machine Learning (ICML).

Health-Informed Policy Gradients for Multi-Agent Reinforcement Learning In International Conference on Machine Learning (ICML)

Reference 2015

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T15:31:05.876700Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-14T15:31:05.448376Z digest=sha256:8d88c02daa122c29a41fead320231b2c2655f08b249d7ec94715889c6685dc25

Observation 181958b0-4945-4f8d-b794-b7216c44b7cf · outbound

This paper cites In International Conference on Learning Representations.

Health-Informed Policy Gradients for Multi-Agent Reinforcement Learning In International Conference on Learning Representations

Reference 2016

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T15:31:05.855426Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-14T15:31:05.457531Z digest=sha256:d2ae25984f5e9ff8ac927bb22caffc68c16dedf82fec32afb59ac649f1272f01

Observation 010840c6-7b1b-4124-a893-4d5449536316 · outbound

This paper cites In Advances in Neural Information Processing Systems (NIPS).

Health-Informed Policy Gradients for Multi-Agent Reinforcement Learning In Advances in Neural Information Processing Systems (NIPS)

Reference 2017

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T15:31:06.001974Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-14T15:31:05.406555Z digest=sha256:409417a2649b6d1d21fccf223254b034e84387b94ab6f5d05babe29fe7bf5b00

Pith citing papers

No inbound Pith citation observations are available.