Pith. sign in

Paper Citation Record · LEDGER

ColorGrid: A Multi-Agent Non-Stationary Environment for Goal Inference and Assistance

As of 13 August 2026, this Paper Citation Record lists 37 of 37 outbound references and 0 inbound Pith citation observations for arXiv:2501.10593.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2501.10593 v1

Coverage vector

measured 37 of 37 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-10T19:06:27.178808Z

measured 37 of 37 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-12T06:34:41.77262+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

37 of 37 outbound references displayed

  • verified exact0
  • verified fuzzy23
  • unresolved14
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation d537b2b4-f8f2-4c17-9e38-338ea14f1ee1 · outbound

This paper cites Basis for intentions: Efficient inverse reinforcement learning using past experience, 2022.

ColorGrid: A Multi-Agent Non-Stationary Environment for Goal Inference and Assistance Basis for intentions: Efficient inverse reinforcement learning using past experience, 2022

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T19:06:29.064232Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T19:06:26.736762Z digest=sha256:73122bcf5510522d5f6aadf6906710fb7658eadc06f8abfcde9358183f90c28d

Observation 7a89f7a3-5005-4e9b-91da-30d964261d5b · outbound

This paper cites Albrecht and Subramanian Ramamoorthy.

ColorGrid: A Multi-Agent Non-Stationary Environment for Goal Inference and Assistance Albrecht and Subramanian Ramamoorthy

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T19:06:29.017281Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T19:06:26.746894Z digest=sha256:cbb74d47d4ebc8b153762df894ddf117cadc1640786a0287358d12ca2ce19309

Observation f3242f02-021c-4f88-961e-d0361f63e8e7 · outbound

This paper cites On the Utility of Learning about Humans for Human-AI Coordination.

ColorGrid: A Multi-Agent Non-Stationary Environment for Goal Inference and Assistance On the Utility of Learning about Humans for Human-AI Coordination

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-10T19:06:26.752581Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T19:06:26.752581Z digest=sha256:0f264dc9d6760b1908dddd194a11e5b3b3ae414b96bbc2b35a585abf1bc788ce

Observation 28190281-ef35-429a-b64f-fe07fe9e79e9 · outbound

This paper cites Is Independent Learning All You Need in the StarCraft Multi-Agent Challenge?.

ColorGrid: A Multi-Agent Non-Stationary Environment for Goal Inference and Assistance Is Independent Learning All You Need in the StarCraft Multi-Agent Challenge?

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-10T19:06:26.778989Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T19:06:26.778989Z digest=sha256:357ddb164d3a9cadea76c958836e5b1de8165668d869b0e7ab9bf7d9e5c54d7f

Observation 4ad79d62-4ff7-4c76-acc4-ab5a830543dd · outbound

This paper cites Emergent complexity and zero-shot transfer via unsupervised environment design, 2021.

ColorGrid: A Multi-Agent Non-Stationary Environment for Goal Inference and Assistance Emergent complexity and zero-shot transfer via unsupervised environment design, 2021

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T19:06:28.978784Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T19:06:26.786278Z digest=sha256:0a9fab8b955fe1005d4c5c9c42169546d74be313db0532e165bd0e9f030dd3af

Observation 696fb80f-8639-428e-8cbd-89b48b18b58c · outbound

This paper cites Counterfactual multi-agent policy gradients, 2017.

ColorGrid: A Multi-Agent Non-Stationary Environment for Goal Inference and Assistance Counterfactual multi-agent policy gradients, 2017

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T19:06:28.943528Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T19:06:26.818117Z digest=sha256:c072bd18bf009ce1f9d9c34ecfd6862cf2cdc7b9b5e2f6441ef1efec5dfeda9a

Observation 18b4baed-7607-44b5-91c4-13a243e35499 · outbound

This paper cites Soft actor-critic: Off-policy maximum entropy deep reinforcement learning with a stochastic actor, 2018.

ColorGrid: A Multi-Agent Non-Stationary Environment for Goal Inference and Assistance Soft actor-critic: Off-policy maximum entropy deep reinforcement learning with a stochastic actor, 2018

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-10T19:06:26.828973Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T19:06:26.828973Z digest=sha256:e82e9ccccbecffea1f35da91c8ad5e54e4b7fee7fe16176acbb70583a5191048

Observation 3694a783-74b9-491e-940f-8dc193a0fc09 · outbound

This paper cites The evolution of cultural evolution.

ColorGrid: A Multi-Agent Non-Stationary Environment for Goal Inference and Assistance The evolution of cultural evolution

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T19:06:28.881357Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T19:06:26.836328Z digest=sha256:68bf987ba45654df065556c963cdfc77bc10168a44ec0b8b52254c6b89d94e71

Observation ba55ccc7-2f1a-4dcd-b82e-cada4a6cc873 · outbound

This paper cites an unresolved cited work.

ColorGrid: A Multi-Agent Non-Stationary Environment for Goal Inference and Assistance Unresolved cited work

Reference 9

Resolution
unresolved
raw_fallback, observed 2026-08-10T19:06:28.836315Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T19:06:26.846414Z digest=sha256:48ec9e99b9ac97d01e9fcc39c675f47159d358846b36ee527de4613043288bb6

Observation 861f1a24-9d73-4110-b6b9-ea501e3081fd · outbound

This paper cites Reinforcement learning with unsupervised auxiliary tasks, 2016.

ColorGrid: A Multi-Agent Non-Stationary Environment for Goal Inference and Assistance Reinforcement learning with unsupervised auxiliary tasks, 2016

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T19:06:28.781669Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T19:06:26.857339Z digest=sha256:d0bcfa7ccce5fcf775954561de7af43f040502d470c936e57528fe5f8c914b52

Observation fb0918b5-983d-4fa2-8d0e-2c38e1180c19 · outbound

This paper cites Czarnecki, Iain Dunning, Luke Marris, Guy Lever, Antonio Garc´ıa Casta˜neda, Charlie Beattie, Neil C.

ColorGrid: A Multi-Agent Non-Stationary Environment for Goal Inference and Assistance Czarnecki, Iain Dunning, Luke Marris, Guy Lever, Antonio Garc´ıa Casta˜neda, Charlie Beattie, Neil C

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T19:06:28.721807Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T19:06:26.868220Z digest=sha256:3213d1a7862c481ea31951287e3f59040bd7ed2b4b89d70472f2723008964c5d

Observation f7a2753b-44eb-4b42-bcdf-5aa763ea501e · outbound

This paper cites Recursive bayesian human intent recognition in shared-control robotics.

ColorGrid: A Multi-Agent Non-Stationary Environment for Goal Inference and Assistance Recursive bayesian human intent recognition in shared-control robotics

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-10T19:06:26.880253Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T19:06:26.880253Z digest=sha256:422413a073874b1dd5af3625eefcf5f9c5a0feb8249dbc851414a8499edbd940

Observation c21e5ce3-3bbc-4023-9cfc-b3d1657398eb · outbound

This paper cites Losey, and Dorsa Sadigh.

ColorGrid: A Multi-Agent Non-Stationary Environment for Goal Inference and Assistance Losey, and Dorsa Sadigh

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T19:06:28.670905Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T19:06:26.888108Z digest=sha256:6ce8dea207ce2ef77f084f5cae67c5f0a466783998d377f92263a50a3b4df719

Observation d077819f-cd94-43bc-98e4-fa3b906f03d8 · outbound

This paper cites Learning dynamics model in reinforcement learning by incorporating the long term future, 2019.

ColorGrid: A Multi-Agent Non-Stationary Environment for Goal Inference and Assistance Learning dynamics model in reinforcement learning by incorporating the long term future, 2019

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T19:06:28.635298Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T19:06:26.905877Z digest=sha256:5b33a3576a2dd35ebd55f6240db5b4b5c3f639c2e82f7c0ff8780c533aea8ceb

Observation 6d49a3f2-990c-4dd7-84ad-83ba94f7f69a · outbound

This paper cites Multi-agent reinforcement learning with multi- step generative models, 2019.

ColorGrid: A Multi-Agent Non-Stationary Environment for Goal Inference and Assistance Multi-agent reinforcement learning with multi- step generative models, 2019

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T19:06:28.598765Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T19:06:26.911966Z digest=sha256:8db501d4b3970241a40c20f77517f4fa6d2bb5110c2351169c27c093e5d70a3c

Observation 6161702c-d913-4fa4-8997-27b3a7f2f943 · outbound

This paper cites Social learning strategies.

ColorGrid: A Multi-Agent Non-Stationary Environment for Goal Inference and Assistance Social learning strategies

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-10T19:06:26.917903Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T19:06:26.917903Z digest=sha256:a33fc039ccd64d923226eae76eaaf42f19395e172cbc59f0fbb112aaf6c2a90f

Observation a56f853e-6062-473d-b088-c8eaa831ca3b · outbound

This paper cites Generalization and network design strategies.

ColorGrid: A Multi-Agent Non-Stationary Environment for Goal Inference and Assistance Generalization and network design strategies

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T19:06:28.508780Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T19:06:26.923958Z digest=sha256:9af0418e09bb1dfd34e9bc92bb24aeaf30f780e5490dd2b231cacb9ba1c62857

Observation 129e91d5-557d-48d8-9146-3c3e908aa58e · outbound

This paper cites Rectifier nonlinearities improve neural network acoustic models.

ColorGrid: A Multi-Agent Non-Stationary Environment for Goal Inference and Assistance Rectifier nonlinearities improve neural network acoustic models

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T19:06:28.453366Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T19:06:26.929333Z digest=sha256:5c9dcf9a476e6f37ef70dbeb87b1ad895ef0b3ceb6ce42b06e62fdac497e914a

Observation c26c8b57-50ca-4be1-b317-1a35c8410603 · outbound

This paper cites Emergence of grounded compositional language in multi- agent populations, 2018.

ColorGrid: A Multi-Agent Non-Stationary Environment for Goal Inference and Assistance Emergence of grounded compositional language in multi- agent populations, 2018

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T19:06:28.370840Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T19:06:26.936560Z digest=sha256:e7d822d9132e59d4c2ccb22e779759695688ec324d42a1e0400fdf11b18110b4

Observation b61c1c8b-7e1b-4ca6-92ba-9a9ae73d3ceb · outbound

This paper cites Emergent social learning via multi-agent reinforcement learning, 2021.

ColorGrid: A Multi-Agent Non-Stationary Environment for Goal Inference and Assistance Emergent social learning via multi-agent reinforcement learning, 2021

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T19:06:28.332017Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T19:06:26.947585Z digest=sha256:db817f346b5c85ebebcb6634bd204942da34d0efbacfec803ed8d768eeb6082f

Observation eeae4ba5-5d80-493a-a368-b540d19c68b1 · outbound

This paper cites Srinivasa.

ColorGrid: A Multi-Agent Non-Stationary Environment for Goal Inference and Assistance Srinivasa

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T19:06:28.296403Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T19:06:26.962941Z digest=sha256:c53352aaac3552f943a96c189e97316037f4c8516a30ae2caf313257f22992e9

Observation f00b01da-f6db-4277-9387-0382f7875662 · outbound

This paper cites Albrecht.

ColorGrid: A Multi-Agent Non-Stationary Environment for Goal Inference and Assistance Albrecht

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T19:06:28.250550Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T19:06:26.982192Z digest=sha256:8141e62bafc706a6bbbc74111ab4015490c0b2d727e29c5805729f4b56ee5357

Observation dc68b850-5a84-4e86-98e7-7f65d2f45bb3 · outbound

This paper cites Tenenbaum, Sanja Fidler, and Antonio Torralba.

ColorGrid: A Multi-Agent Non-Stationary Environment for Goal Inference and Assistance Tenenbaum, Sanja Fidler, and Antonio Torralba

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T19:06:28.181958Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T19:06:27.011753Z digest=sha256:1dadf1222f9bad9cf9070330e9b0aadea5776d627be1df73eda5a7d2bf4e83e7

Observation b08e9966-4202-4252-be82-8c096ead850a · outbound

This paper cites Modeling others using oneself in multi-agent reinforcement learning, 2018.

ColorGrid: A Multi-Agent Non-Stationary Environment for Goal Inference and Assistance Modeling others using oneself in multi-agent reinforcement learning, 2018

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T19:06:28.125248Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T19:06:27.026167Z digest=sha256:d18c37098e49e9096f5bae11b66ba81057e2f1c1b7334033ac6d17e113f66993

Observation 2673280f-6d84-4a65-a09b-c6cee5f12feb · outbound

This paper cites Proximal policy optimization algorithms, 2017.

ColorGrid: A Multi-Agent Non-Stationary Environment for Goal Inference and Assistance Proximal policy optimization algorithms, 2017

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-10T19:06:27.034876Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T19:06:27.034876Z digest=sha256:fbb11b25edb5d0ac0481829f096b3b6dc4ceddff6b46be7d817488f3cca69bb7

Observation af535b3f-c668-4f6e-a907-9bba527059b1 · outbound

This paper cites Loss is its own reward: Self-supervision for reinforcement learning, 2017.

ColorGrid: A Multi-Agent Non-Stationary Environment for Goal Inference and Assistance Loss is its own reward: Self-supervision for reinforcement learning, 2017

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T19:06:27.994120Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T19:06:27.046728Z digest=sha256:6a8628ec467a5016087e3eb1887b491501153e20423b61bcb8ea149b6dc6c58f

Observation 96d206e5-8946-4903-8469-b644aa0c6ee6 · outbound

This paper cites Zhao, Archit Sharma, Karl Pertsch, Jianlan Luo, Sergey Levine, and Chelsea Finn.

ColorGrid: A Multi-Agent Non-Stationary Environment for Goal Inference and Assistance Zhao, Archit Sharma, Karl Pertsch, Jianlan Luo, Sergey Levine, and Chelsea Finn

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-10T19:06:27.058046Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T19:06:27.058046Z digest=sha256:caaf98a2928cd2b335ae89379a970699e2be8967c32ec031b25c34987b0f764d

Observation 1cbbcdb1-cd8c-42ef-b820-3a27a3bb2390 · outbound

This paper cites Message-passing approach for threshold models of behavior in networks.

ColorGrid: A Multi-Agent Non-Stationary Environment for Goal Inference and Assistance Message-passing approach for threshold models of behavior in networks

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-10T19:06:27.072016Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T19:06:27.072016Z digest=sha256:3f5f62850533d1b1c3bb5a65de350c1fc1998f690ac25e6dde7a22b2afc723fb

Observation 4068f696-d52e-414c-b6f1-74108fd1f5e1 · outbound

This paper cites Mastering chess and shogi by self-play with a general reinforcement learning algorithm, 2017.

ColorGrid: A Multi-Agent Non-Stationary Environment for Goal Inference and Assistance Mastering chess and shogi by self-play with a general reinforcement learning algorithm, 2017

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-10T19:06:27.080776Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T19:06:27.080776Z digest=sha256:d975af67715a1abf8ab0d27eb8004afb208cf067171fcee1875f949ef06eb293

Observation 0e5e48d1-68a2-4b52-bbe4-7563b449296d · outbound

This paper cites Smallwood and Edward J.

ColorGrid: A Multi-Agent Non-Stationary Environment for Goal Inference and Assistance Smallwood and Edward J

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T19:06:27.918249Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T19:06:27.094921Z digest=sha256:072f1c849820192244668624cc8990af60c91549815d6fc89a91f405db1cd4d0

Observation 6d44c7e6-4d02-4634-baa4-49da968a5114 · outbound

This paper cites McKee, Matt Botvinick, Edward Hughes, and Richard Everett.

ColorGrid: A Multi-Agent Non-Stationary Environment for Goal Inference and Assistance McKee, Matt Botvinick, Edward Hughes, and Richard Everett

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T19:06:27.878794Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T19:06:27.103156Z digest=sha256:06ada664174e20d7b2f832b38f0546787e306c46da176398aab90b9741312055

Observation be4d5c22-43f9-46aa-8646-0eff70c945bb · outbound

This paper cites Pettingzoo: Gym for multi-agent reinforcement learning.

ColorGrid: A Multi-Agent Non-Stationary Environment for Goal Inference and Assistance Pettingzoo: Gym for multi-agent reinforcement learning

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-10T19:06:27.109331Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T19:06:27.109331Z digest=sha256:5a02a0f3b5987d0075ee7cd68c117063053290561e97f640c5d2fc2301706e69

Observation fe531969-7b1a-44a8-9409-6e6f0018e467 · outbound

This paper cites Paired Open-Ended Trailblazer (POET): Endlessly Generating Increasingly Complex and Diverse Learning Environments and Their Solutions.

ColorGrid: A Multi-Agent Non-Stationary Environment for Goal Inference and Assistance Paired Open-Ended Trailblazer (POET): Endlessly Generating Increasingly Complex and Diverse Learning Environments and Their Solutions

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-10T19:06:27.125935Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T19:06:27.125935Z digest=sha256:52a3809f08fb7279f91215319d8acd0c70de33092b6bf59e089eb7c830f151b3

Observation dc7e7e4a-e1b2-4294-bcfd-d658448eb67e · outbound

This paper cites an unresolved cited work.

ColorGrid: A Multi-Agent Non-Stationary Environment for Goal Inference and Assistance Unresolved cited work

Reference 34

Resolution
unresolved
raw_fallback, observed 2026-08-10T19:06:27.831658Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T19:06:27.151138Z digest=sha256:5a24b51b04ea1f38e2d0c5fa680131acdbaef0a7db684833f225d36a9ace133a

Observation 10948f30-246e-4d58-bc53-08f5d4e22f80 · outbound

This paper cites Towards generalizability of multi-agent reinforcement learning in graphs with recurrent message passing, 2024.

ColorGrid: A Multi-Agent Non-Stationary Environment for Goal Inference and Assistance Towards generalizability of multi-agent reinforcement learning in graphs with recurrent message passing, 2024

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T19:06:27.782180Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T19:06:27.166033Z digest=sha256:65386cc251cb4d1dbcbc299ec49e34790ce7b4bca33a6f8c6fd6085272e5580d

Observation 12201486-51f9-4452-a6b7-4276c427015c · outbound

This paper cites The surprising effectiveness of ppo in cooperative, multi-agent games, 2021.

ColorGrid: A Multi-Agent Non-Stationary Environment for Goal Inference and Assistance The surprising effectiveness of ppo in cooperative, multi-agent games, 2021

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T19:06:27.717210Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T19:06:27.178808Z digest=sha256:c9e248b41a09af075909aeb767a08c538d18343e07346b0afb76ef215e4ab987

Observation 58a30f00-1af9-4bc3-bc2e-861941e7d3c0 · outbound

This paper cites Benchmarking Multi-Agent Deep Reinforcement Learning Algorithms in Cooperative Tasks.

ColorGrid: A Multi-Agent Non-Stationary Environment for Goal Inference and Assistance Benchmarking Multi-Agent Deep Reinforcement Learning Algorithms in Cooperative Tasks

Reference 2021

Resolution
unresolved
no resolver link, observed 2026-08-10T19:06:26.997151Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T19:06:26.997151Z digest=sha256:a4459d00d44f141bfa82c20d6b7d9ef2ef2484d3541315224bfa1b0a7c279be0

Pith citing papers

No inbound Pith citation observations are available.