Pith. sign in

Paper Citation Record · LEDGER

Health-Informed Policy Gradients for Multi-Agent Reinforcement Learning

As of 20 August 2026, this Paper Citation Record lists 43 of 43 outbound references and 0 inbound Pith citation observations for arXiv:1908.01022.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
1908.01022 v4

Coverage vector

measured 43 of 43 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-14T15:31:05.521816Z

measured 43 of 43 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-20T06:33:59.587034+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

43 of 43 outbound references displayed

  • verified exact1
  • verified fuzzy4
  • unresolved38
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 65b0b90a-3e02-49e0-95da-dcb52ce62544 · outbound

This paper cites Johnson, and Mykel J.

Health-Informed Policy Gradients for Multi-Agent Reinforcement Learning Johnson, and Mykel J

Reference 1

Resolution
verified exact
doi, observed 2026-08-14T15:31:05.566753Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-14T15:31:05.310320Z digest=sha256:f4a6df8554ffb0e5f26c32a4de7f745a7e97ae12a24a987ebb78a24d138a1025

Observation 2abd25d9-1d41-4b62-b7a1-cba858270789 · outbound

This paper cites an unresolved cited work.

Health-Informed Policy Gradients for Multi-Agent Reinforcement Learning Unresolved cited work

Reference 2

Resolution
unresolved
raw_fallback, observed 2026-08-14T15:31:06.175089Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-14T15:31:05.316614Z digest=sha256:beeb86677e3633ce9f8371372c24e0764929c86ab946cd87cdb51e59ac15641d

Observation c3f3bea2-e71a-45a9-b433-67b050cac5b1 · outbound

This paper cites an unresolved cited work.

Health-Informed Policy Gradients for Multi-Agent Reinforcement Learning Unresolved cited work

Reference 3

Resolution
unresolved
raw_fallback, observed 2026-08-14T15:31:06.162235Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-14T15:31:05.323044Z digest=sha256:9e170f12aea47e057a5609ff48ec16580bc998649da023fe4ce6d35446943726

Observation b11fd6d3-dd02-4198-b51f-42c40520d513 · outbound

This paper cites an unresolved cited work.

Health-Informed Policy Gradients for Multi-Agent Reinforcement Learning Unresolved cited work

Reference 4

Resolution
unresolved
raw_fallback, observed 2026-08-14T15:31:06.148539Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-14T15:31:05.328023Z digest=sha256:be2129f3a94fa471de5ae5da80fd4c82667179b6e050f117f9928c44b510504a

Observation 9e36de8c-4caf-4d17-8afe-26f0fa00b6f9 · outbound

This paper cites Fast and Accurate Deep Network Learning by Exponential Linear Units (ELUs).

Health-Informed Policy Gradients for Multi-Agent Reinforcement Learning Fast and Accurate Deep Network Learning by Exponential Linear Units (ELUs)

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-14T15:31:05.334928Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T15:31:05.334928Z digest=sha256:f2b9d4ec908241c08b931fe78ffef0c40e9a382dd62734ac234172e5e1999ff3

Observation 390e912c-f17c-4e01-83da-bf9735861889 · outbound

This paper cites an unresolved cited work.

Health-Informed Policy Gradients for Multi-Agent Reinforcement Learning Unresolved cited work

Reference 6

Resolution
unresolved
raw_fallback, observed 2026-08-14T15:31:06.134754Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-14T15:31:05.340214Z digest=sha256:f255c18e952d916d7a9dd0505c975b4a41a6da18a8f856ffca9c17cbc1472abf

Observation 31080e64-0030-4af8-9554-dd8f96dbe302 · outbound

This paper cites an unresolved cited work.

Health-Informed Policy Gradients for Multi-Agent Reinforcement Learning Unresolved cited work

Reference 7

Resolution
unresolved
raw_fallback, observed 2026-08-14T15:31:06.120944Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-14T15:31:05.346065Z digest=sha256:775ea6ad8cdf5187d11db6b3635f6b5ba1ef578bf4c3afdeb4c027afc7df81e3

Observation cbcaafa7-d455-418d-8f95-a939dcbc5237 · outbound

This paper cites Clipped Action Policy Gradient.

Health-Informed Policy Gradients for Multi-Agent Reinforcement Learning Clipped Action Policy Gradient

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-14T15:31:05.351418Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T15:31:05.351418Z digest=sha256:9fdec3f0d787a3ecc54cc8b189ded2f2e8ac34cab6c0ad0af59d87b413fd8c48

Observation a7b9d074-431a-4ed1-9a19-397a40aca2d9 · outbound

This paper cites an unresolved cited work.

Health-Informed Policy Gradients for Multi-Agent Reinforcement Learning Unresolved cited work

Reference 9

Resolution
unresolved
raw_fallback, observed 2026-08-14T15:31:06.107449Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-14T15:31:05.358713Z digest=sha256:35ec65a7968d96aab845f8fa6447043011671cf7598ed382a71aace770e4ff09

Observation 77bfb7dd-02fe-43fc-aa02-000d6e612255 · outbound

This paper cites an unresolved cited work.

Health-Informed Policy Gradients for Multi-Agent Reinforcement Learning Unresolved cited work

Reference 10

Resolution
unresolved
raw_fallback, observed 2026-08-14T15:31:06.093176Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-14T15:31:05.364252Z digest=sha256:2382919e43980443fefc3af359e215253854229c7aff7c7ba049ac56245266fd

Observation 95fa0923-ad83-4470-9988-dab739bc888e · outbound

This paper cites A Survey of Learning in Multiagent Environments: Dealing with Non-Stationarity.

Health-Informed Policy Gradients for Multi-Agent Reinforcement Learning A Survey of Learning in Multiagent Environments: Dealing with Non-Stationarity

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-14T15:31:05.371369Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T15:31:05.371369Z digest=sha256:901cb35f53d6eb430ce22ddb0d1dd8776d247c07424e60db75e6eab67765fe3e

Observation 32cf939a-3c82-4e49-8946-0fa8e015955c · outbound

This paper cites an unresolved cited work.

Health-Informed Policy Gradients for Multi-Agent Reinforcement Learning Unresolved cited work

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-14T15:31:05.375880Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T15:31:05.375880Z digest=sha256:19ebf56a0d9020cae90a532e2ab60760947361194dafdda99f8819d30956107b

Observation 8f3f02ff-79d9-4035-8c09-03ce77149e4e · outbound

This paper cites an unresolved cited work.

Health-Informed Policy Gradients for Multi-Agent Reinforcement Learning Unresolved cited work

Reference 13

Resolution
unresolved
raw_fallback, observed 2026-08-14T15:31:06.067606Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-14T15:31:05.380130Z digest=sha256:b70a715f5ce6264bbe34d7c028ebda453f4c0efcec63715520f21240abe3d9cd

Observation 32931c9f-8002-4bc8-8fd4-ebd9481ddb33 · outbound

This paper cites Adam: A Method for Stochastic Optimization.

Health-Informed Policy Gradients for Multi-Agent Reinforcement Learning Adam: A Method for Stochastic Optimization

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-14T15:31:05.384134Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T15:31:05.384134Z digest=sha256:f14d982996032899461d643213112abaf57a1d9517355ff9a5e45ccb7f2b80e1

Observation 3720ba56-e903-4da3-bbfe-a3d0ea01153d · outbound

This paper cites an unresolved cited work.

Health-Informed Policy Gradients for Multi-Agent Reinforcement Learning Unresolved cited work

Reference 15

Resolution
unresolved
raw_fallback, observed 2026-08-14T15:31:06.050520Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-14T15:31:05.388898Z digest=sha256:8d2d0d1c9363195de2f040a5e23cf5b3210373eee02732153a16fe421f320ef2

Observation 367d1e24-2482-48c5-a289-cfcb8691f10a · outbound

This paper cites an unresolved cited work.

Health-Informed Policy Gradients for Multi-Agent Reinforcement Learning Unresolved cited work

Reference 16

Resolution
unresolved
raw_fallback, observed 2026-08-14T15:31:06.035364Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-14T15:31:05.392726Z digest=sha256:9a095f44e9a059c5d5f0c44c37f59d16dd45daac91a5cd5822e9b2f29505844e

Observation e6754999-4130-47a4-b56f-fe4a4e9ab889 · outbound

This paper cites an unresolved cited work.

Health-Informed Policy Gradients for Multi-Agent Reinforcement Learning Unresolved cited work

Reference 17

Resolution
unresolved
raw_fallback, observed 2026-08-14T15:31:06.022486Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-14T15:31:05.396605Z digest=sha256:ffd3c537f5b92d2dbf6d1e233eb3f2df8e5794197805fa73ccb7d0634425be29

Observation 25514841-e649-49eb-a425-5ff16000293e · outbound

This paper cites an unresolved cited work.

Health-Informed Policy Gradients for Multi-Agent Reinforcement Learning Unresolved cited work

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-14T15:31:05.402129Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T15:31:05.402129Z digest=sha256:c5876809734008e67bdfd0e9404885ecc350d31343e97514724d4989a7040bc8

Observation 317920fb-7582-45ea-9fc8-1f12ddfbb102 · outbound

This paper cites an unresolved cited work.

Health-Informed Policy Gradients for Multi-Agent Reinforcement Learning Unresolved cited work

Reference 19

Resolution
unresolved
raw_fallback, observed 2026-08-14T15:31:05.987246Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-14T15:31:05.411482Z digest=sha256:0cc81d2c32e38d4279d403e194fc9d7873cf4c66ebe67843afe75c35b504dd62

Observation 2515a4c3-9227-4b91-af23-521c1ba2ff2c · outbound

This paper cites an unresolved cited work.

Health-Informed Policy Gradients for Multi-Agent Reinforcement Learning Unresolved cited work

Reference 20

Resolution
unresolved
raw_fallback, observed 2026-08-14T15:31:05.973456Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-14T15:31:05.415148Z digest=sha256:a9241bd894528dd08d97b8e414c3f945503bfb397cd2f4663e4c21623742d586

Observation fff15ea5-9d5e-40f3-be5d-e30aadbc7c29 · outbound

This paper cites an unresolved cited work.

Health-Informed Policy Gradients for Multi-Agent Reinforcement Learning Unresolved cited work

Reference 21

Resolution
unresolved
raw_fallback, observed 2026-08-14T15:31:05.957429Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-14T15:31:05.418822Z digest=sha256:791ed865e06ff07a5881ca8402e9a7a213a35bf4bc67979516d86379ca0218f0

Observation a8344aac-f8b1-42e0-9db0-9a715a16d13f · outbound

This paper cites an unresolved cited work.

Health-Informed Policy Gradients for Multi-Agent Reinforcement Learning Unresolved cited work

Reference 22

Resolution
unresolved
raw_fallback, observed 2026-08-14T15:31:05.944517Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-14T15:31:05.422232Z digest=sha256:be9d1ef0a921acb7d183c2d3ed7b43053117ce73d3e8474dfb14852e0e3b0bdb

Observation 89d0028c-8ae7-42f3-9dae-0dd1ba801994 · outbound

This paper cites an unresolved cited work.

Health-Informed Policy Gradients for Multi-Agent Reinforcement Learning Unresolved cited work

Reference 23

Resolution
unresolved
raw_fallback, observed 2026-08-14T15:31:05.929995Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-14T15:31:05.426347Z digest=sha256:3650f2a812880b4d8b823daa662afc7ed8828ee5c5be1e4f03cd7c1ad7f30064

Observation f9d723fa-679b-4101-ba09-bc228dbe0f23 · outbound

This paper cites an unresolved cited work.

Health-Informed Policy Gradients for Multi-Agent Reinforcement Learning Unresolved cited work

Reference 24

Resolution
unresolved
raw_fallback, observed 2026-08-14T15:31:05.916570Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-14T15:31:05.429813Z digest=sha256:84ee13075d046309cb556cf55f951428c514197675224eec40143454315b5170

Observation d80faefe-d94e-4247-907e-f12844b09d40 · outbound

This paper cites QMIX: Monotonic Value Function Factorisation for Deep Multi-Agent Reinforcement Learning.

Health-Informed Policy Gradients for Multi-Agent Reinforcement Learning QMIX: Monotonic Value Function Factorisation for Deep Multi-Agent Reinforcement Learning

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-14T15:31:05.435396Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T15:31:05.435396Z digest=sha256:7a7ed189093f8a45d92a2d40a9ab20047ba8415bae186adc044dcb80a4e1bf7a

Observation cee04415-5860-4db5-a79d-4f347956314c · outbound

This paper cites an unresolved cited work.

Health-Informed Policy Gradients for Multi-Agent Reinforcement Learning Unresolved cited work

Reference 26

Resolution
unresolved
raw_fallback, observed 2026-08-14T15:31:05.901447Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-14T15:31:05.439753Z digest=sha256:e26a6d50c86ead8b4db1c1b2980304e114da9eff33e9483721782200a0fa69b2

Observation c12ff084-751d-4dd2-84a8-2d52dcc6b8e2 · outbound

This paper cites an unresolved cited work.

Health-Informed Policy Gradients for Multi-Agent Reinforcement Learning Unresolved cited work

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-14T15:31:05.443913Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T15:31:05.443913Z digest=sha256:60bd07a1379b981d219e7e540a8f0b022b6abdcdfde7032c547fea484f7c02cd

Observation 1c33cf1f-9f28-4881-81d1-1df3056d797e · outbound

This paper cites an unresolved cited work.

Health-Informed Policy Gradients for Multi-Agent Reinforcement Learning Unresolved cited work

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-14T15:31:05.452852Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T15:31:05.452852Z digest=sha256:c5bd6f690ecf8eed349c6f8c060aaa4701d9592c0d62970c5a28373f16f182af

Observation 3a596216-928c-4a0b-9e78-04139c97f911 · outbound

This paper cites an unresolved cited work.

Health-Informed Policy Gradients for Multi-Agent Reinforcement Learning Unresolved cited work

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-14T15:31:05.461962Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T15:31:05.461962Z digest=sha256:80141962f49d8ce8df96066beeea37704b836325cc0e05df147b58ab9d98a9a6

Observation 56e5fd1e-a936-41b0-8b26-a7f7442ea947 · outbound

This paper cites an unresolved cited work.

Health-Informed Policy Gradients for Multi-Agent Reinforcement Learning Unresolved cited work

Reference 30

Resolution
unresolved
raw_fallback, observed 2026-08-14T15:31:05.818430Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-14T15:31:05.471601Z digest=sha256:3ed1cd0cd558bab547e4ed91d2c90ae28eb11713a2f48a9d7c7858da043342eb

Observation 86865a61-7c71-4f0d-b72a-c520e833a3cc · outbound

This paper cites an unresolved cited work.

Health-Informed Policy Gradients for Multi-Agent Reinforcement Learning Unresolved cited work

Reference 31

Resolution
unresolved
raw_fallback, observed 2026-08-14T15:31:05.805674Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-14T15:31:05.476887Z digest=sha256:e2ff2651df017494eae25a58daa31f4c7b5757b68886606bf3410d16c36097c8

Observation ef9503ff-9247-442c-9534-86bfd7386253 · outbound

This paper cites an unresolved cited work.

Health-Informed Policy Gradients for Multi-Agent Reinforcement Learning Unresolved cited work

Reference 32

Resolution
unresolved
raw_fallback, observed 2026-08-14T15:31:05.790913Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-14T15:31:05.481402Z digest=sha256:82efe94ed4e266490a63426a47116a27dc5384b5ec8c0492ec4be2af0edb9959

Observation 01eef5a0-2d6b-4956-b49b-91a666d4146e · outbound

This paper cites Proximal Policy Optimization Algorithms.

Health-Informed Policy Gradients for Multi-Agent Reinforcement Learning Proximal Policy Optimization Algorithms

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-14T15:31:05.466656Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T15:31:05.466656Z digest=sha256:d3a04775eeb1514fb929916f14c70a9dd8a958ca0d954b725a57906bdde9d5bb

Observation 94e8f63d-5a69-4f53-b82c-520e8c7c1adb · outbound

This paper cites PettingZoo: Gym for Multi-Agent Reinforcement Learning.

Health-Informed Policy Gradients for Multi-Agent Reinforcement Learning PettingZoo: Gym for Multi-Agent Reinforcement Learning

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-14T15:31:05.493756Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T15:31:05.493756Z digest=sha256:24f32324ce4fcff10ae6e8b836a9439827a5c8b232c22b32804704788c1434fd

Observation 1369fbe1-dc29-407e-846b-67a129431d07 · outbound

This paper cites Revisiting Parameter Sharing in Multi-Agent Deep Reinforcement Learning.

Health-Informed Policy Gradients for Multi-Agent Reinforcement Learning Revisiting Parameter Sharing in Multi-Agent Deep Reinforcement Learning

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-14T15:31:05.499450Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T15:31:05.499450Z digest=sha256:a8a87988c16ef35deff250aca81a57e7c923a35c2585e7f4048a45b68874041e

Observation d17a4c03-1a2a-4d9f-b905-e33d00b3d7a6 · outbound

This paper cites an unresolved cited work.

Health-Informed Policy Gradients for Multi-Agent Reinforcement Learning Unresolved cited work

Reference 36

Resolution
unresolved
raw_fallback, observed 2026-08-14T15:31:05.754955Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-14T15:31:05.506160Z digest=sha256:898e33276fefbab334df29b6eac89011c3fec1ac3956f8caa61bafa2fc26ae9d

Observation 9ddddd13-3209-49eb-ac6d-e6f38885c044 · outbound

This paper cites an unresolved cited work.

Health-Informed Policy Gradients for Multi-Agent Reinforcement Learning Unresolved cited work

Reference 37

Resolution
unresolved
raw_fallback, observed 2026-08-14T15:31:05.772770Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-14T15:31:05.487612Z digest=sha256:72436c6b775cce5c4b62c3c350cfe60ceb194749f1d70bf38b374589cea1abe1

Observation 41b53b4a-d3e6-475f-97e7-f9b154727802 · outbound

This paper cites an unresolved cited work.

Health-Informed Policy Gradients for Multi-Agent Reinforcement Learning Unresolved cited work

Reference 38

Resolution
unresolved
raw_fallback, observed 2026-08-14T15:31:05.728186Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-14T15:31:05.516684Z digest=sha256:274966895f18bc68eafdbfbf354484595274382b43be8e1d51b15f12951b5c6b

Observation 2709a228-c544-464b-bf22-7d02fea8d23b · outbound

This paper cites ∑︁ u π(u| s𝑡, θ)· 𝑛∑︁ 𝑖=1 𝑞π(s𝑡, u)∇𝜃𝑖 log𝜋𝑖(𝑎𝑖|𝜏𝑖,𝑡,𝜃𝑖) # = Eπ.

Health-Informed Policy Gradients for Multi-Agent Reinforcement Learning ∑︁ u π(u| s𝑡, θ)· 𝑛∑︁ 𝑖=1 𝑞π(s𝑡, u)∇𝜃𝑖 log𝜋𝑖(𝑎𝑖|𝜏𝑖,𝑡,𝜃𝑖) # = Eπ

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T15:31:05.710236Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-14T15:31:05.521816Z digest=sha256:d6f22e3c2989f3ab58f1ce4d4b4d6d8dbdb7e145d264314db56bca75d0c35b3c

Observation 9a3b754f-7ca8-43d6-b5d7-bae40ef791ec · outbound

This paper cites an unresolved cited work.

Health-Informed Policy Gradients for Multi-Agent Reinforcement Learning Unresolved cited work

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-14T15:31:05.510597Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T15:31:05.510597Z digest=sha256:ff847d90ebcad95262f58c8d68eaa26060ed8dfa084b076e151d483a5f1abc44

Observation 7d9c4ee0-fbdf-4f9c-87c5-9a2b2759cfc8 · outbound

This paper cites In International Conference on Machine Learning (ICML).

Health-Informed Policy Gradients for Multi-Agent Reinforcement Learning In International Conference on Machine Learning (ICML)

Reference 2015

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T15:31:05.876700Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-14T15:31:05.448376Z digest=sha256:b41b4c79cbc4505be545d5cb7ec3ee21beb743f9ed30c073f80de7e6e46db25c

Observation 181958b0-4945-4f8d-b794-b7216c44b7cf · outbound

This paper cites In International Conference on Learning Representations.

Health-Informed Policy Gradients for Multi-Agent Reinforcement Learning In International Conference on Learning Representations

Reference 2016

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T15:31:05.855426Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-14T15:31:05.457531Z digest=sha256:ea8d0e8403c2f5011b1c35bbec6819a4b87537ce858e022b864e5ca6a0999767

Observation 010840c6-7b1b-4124-a893-4d5449536316 · outbound

This paper cites In Advances in Neural Information Processing Systems (NIPS).

Health-Informed Policy Gradients for Multi-Agent Reinforcement Learning In Advances in Neural Information Processing Systems (NIPS)

Reference 2017

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T15:31:06.001974Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-14T15:31:05.406555Z digest=sha256:6cb9afac899bba5a76c79c4851ab6fb1bd43250915737be2acab5b00b4bbefca

Pith citing papers

No inbound Pith citation observations are available.