Pith. sign in

Paper Citation Record · LEDGER

Remembering the Markov Property in Cooperative MARL

As of 17 August 2026, this Paper Citation Record lists 36 of 36 outbound references and 0 inbound Pith citation observations for arXiv:2507.18333.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2507.18333 v1

Coverage vector

measured 36 of 36 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-15T18:20:45.426207Z

measured 36 of 36 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-17T06:30:58.91139+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

36 of 36 outbound references displayed

  • verified exact0
  • verified fuzzy18
  • unresolved18
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation c4fa5e91-f1d6-45c6-86fc-945723ea4e3a · outbound

This paper cites Autonomous agents modelling other agents: A comprehensive survey and open problems.

Remembering the Markov Property in Cooperative MARL Autonomous agents modelling other agents: A comprehensive survey and open problems

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T18:20:46.570620Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-15T18:20:45.070848Z digest=sha256:8726e9ad8ed947d6f9a04a941caf6a1e06506b66cb40bda25778de185cdceadc

Observation 968dbb79-5f4f-43af-9998-6769f5324cf0 · outbound

This paper cites Albrecht, Filippos Christianos, and Lukas Sch\"afer.

Remembering the Markov Property in Cooperative MARL Albrecht, Filippos Christianos, and Lukas Sch\"afer

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T18:20:46.537761Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-15T18:20:45.083380Z digest=sha256:d1bb1466724c5c6161480a78a9d1cd7e0819d69318ed0bd1d3627601aa3701ee

Observation 540b63bb-d6bb-4ac2-ac8c-5c86412e4a62 · outbound

This paper cites Optimal control of M arkov processes with incomplete state information.

Remembering the Markov Property in Cooperative MARL Optimal control of M arkov processes with incomplete state information

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T18:20:46.509296Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-15T18:20:45.093823Z digest=sha256:ea1fcface775357a8289008cce051b2adbb2b47671ad89382d3a9f8d986d6657

Observation 40d5592a-1926-4508-b5b6-acd85a12b731 · outbound

This paper cites The hanabi challenge: A new frontier for ai research.

Remembering the Markov Property in Cooperative MARL The hanabi challenge: A new frontier for ai research

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-15T18:20:45.098662Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T18:20:45.098662Z digest=sha256:c45822d9316d44cb07768c1d9b006f47ecf00e468adf69c2f22dad27577ca98f

Observation 44b23c26-441a-41ba-bf1c-212456047777 · outbound

This paper cites The complexity of decentralized control of markov decision processes.

Remembering the Markov Property in Cooperative MARL The complexity of decentralized control of markov decision processes

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T18:20:46.444200Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-15T18:20:45.105513Z digest=sha256:75e0fcb3958b41b4a9cea256cdc21486cdaea6d703caa9300cc79d5105113bdf

Observation 1bdcf2c9-902d-456a-b0ff-2098cb75461e · outbound

This paper cites Is Independent Learning All You Need in the StarCraft Multi-Agent Challenge?.

Remembering the Markov Property in Cooperative MARL Is Independent Learning All You Need in the StarCraft Multi-Agent Challenge?

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-15T18:20:45.124963Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T18:20:45.124963Z digest=sha256:4693fbb0aeec71eb012937b99051c700825677922982b0fa9dfe299a52d9a336

Observation 767b0a5c-831e-4cfd-8332-42004233aeec · outbound

This paper cites Smacv2: An improved benchmark for cooperative multi-agent reinforcement learning.

Remembering the Markov Property in Cooperative MARL Smacv2: An improved benchmark for cooperative multi-agent reinforcement learning

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T18:20:46.406440Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-15T18:20:45.134830Z digest=sha256:dfe57606852e3c79d845827e6d634133947790c9ebac54ce1980fb4fa29cc924

Observation 9318dc74-91e6-4a2c-989f-9a8ce814a138 · outbound

This paper cites Bayesian action decoder for deep multi-agent reinforcement learning.

Remembering the Markov Property in Cooperative MARL Bayesian action decoder for deep multi-agent reinforcement learning

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T18:20:46.376813Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-15T18:20:45.142457Z digest=sha256:fde2dcdee2506957da9d1927b0197411619e8917c10afc6ff22bf9da252f9576

Observation 9803b937-1ae1-49e1-b051-92f459b39cb4 · outbound

This paper cites Studying the Interplay Between the Actor and Critic Representations in Reinforcement Learning.

Remembering the Markov Property in Cooperative MARL Studying the Interplay Between the Actor and Critic Representations in Reinforcement Learning

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-15T18:20:45.151102Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T18:20:45.151102Z digest=sha256:074f898cf1a258561ff631f1c764f9927bc3c2e6887df5a9f8947361e0a58598

Observation 9e07b5b1-b7f7-478c-b092-f0762c3bc2e0 · outbound

This paper cites Deep recurrent q-learning for partially observable mdps.

Remembering the Markov Property in Cooperative MARL Deep recurrent q-learning for partially observable mdps

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T18:20:46.340119Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-15T18:20:45.158224Z digest=sha256:5d1385b38c8533dc5ace3c2c453e6b7d8119a5d88cf01942232b826a87ba3789

Observation 508e3ad0-4800-4341-9f31-13c702551081 · outbound

This paper cites other-play.

Remembering the Markov Property in Cooperative MARL other-play

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T18:20:46.302156Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-15T18:20:45.169641Z digest=sha256:fb51a6d11ade0c8444188b440862038c75483fb3742c5baea3818e8ef7f85fe5

Observation 961d6cde-5353-416b-91b5-f694ebd03589 · outbound

This paper cites Off-belief learning.

Remembering the Markov Property in Cooperative MARL Off-belief learning

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T18:20:46.257877Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-15T18:20:45.179013Z digest=sha256:ec422521bcc2aa16500a4a5d48ab8eaeac88215993015d1b58ead6ca7e6af7de

Observation d6b2a152-6b39-48f7-8bab-7b0d194a7090 · outbound

This paper cites Planning and acting in partially observable stochastic domains.

Remembering the Markov Property in Cooperative MARL Planning and acting in partially observable stochastic domains

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-15T18:20:45.185723Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T18:20:45.185723Z digest=sha256:6202c2362146a2caa80d5f0429d7f952dfd7c63150ef47d3467b63dfc5ea4838

Observation b75f4662-c49c-44a7-af88-7a7ded0a3d0f · outbound

This paper cites Multi-agent reinforcement learning as a rehearsal for decentralized planning.

Remembering the Markov Property in Cooperative MARL Multi-agent reinforcement learning as a rehearsal for decentralized planning

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-15T18:20:45.191290Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T18:20:45.191290Z digest=sha256:046f2b716c4dad6d2917d6e5f69c79834ee867788f165e6e4b34f65f01dfeafb

Observation 7c522fe2-f1e7-4966-a875-f02cf11c1cc1 · outbound

This paper cites Estimating mutual information.

Remembering the Markov Property in Cooperative MARL Estimating mutual information

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-15T18:20:45.204100Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T18:20:45.204100Z digest=sha256:885327b8b78f8b6ed6682e0f4fb3ad72166891aa9a70c16a300970edd940deff

Observation b35c1ed4-a17e-487d-bda5-bd69207728b4 · outbound

This paper cites Nonapproximability results for partially observable M arkov D ecision P rocesses.

Remembering the Markov Property in Cooperative MARL Nonapproximability results for partially observable M arkov D ecision P rocesses

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T18:20:46.152383Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-15T18:20:45.212871Z digest=sha256:80dd7c496036531f5e59e1ce9de1e3c8c178c250ae273e6a1485b30f8287f6b7

Observation 65ff5392-b109-466f-9f28-5830e7754f26 · outbound

This paper cites Partner Modelling Emerges in Recurrent Agents (But Only When It Matters).

Remembering the Markov Property in Cooperative MARL Partner Modelling Emerges in Recurrent Agents (But Only When It Matters)

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-15T18:20:45.219324Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T18:20:45.219324Z digest=sha256:eabfb3e6e1b934fd770063e650792cb380d9e1954a03722035fb7c0befd60fb1

Observation 62971bf3-cf82-4957-b1ff-1aa10fd3830f · outbound

This paper cites Towards few-shot coordination: Revisiting ad-hoc teamplay challenge in the game of hanabi.

Remembering the Markov Property in Cooperative MARL Towards few-shot coordination: Revisiting ad-hoc teamplay challenge in the game of hanabi

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T18:20:46.100610Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-15T18:20:45.226642Z digest=sha256:c7411c4dca03bc5411e6db02274d025f83b6b3171bb351793d07f35e59bc86ce

Observation cdacd85c-232d-42b7-834e-ee46f00fa7ad · outbound

This paper cites Optimal and approximate q-value functions for decentralized pomdps.

Remembering the Markov Property in Cooperative MARL Optimal and approximate q-value functions for decentralized pomdps

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-15T18:20:45.236391Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T18:20:45.236391Z digest=sha256:0ffc997d04432574cb139d16d3ae7ddedfcaf6f79339711f0287c96885875e30

Observation 163e3c6b-7ab6-4758-b4c4-75978526783a · outbound

This paper cites A concise introduction to decentralized POMDPs, volume 1.

Remembering the Markov Property in Cooperative MARL A concise introduction to decentralized POMDPs, volume 1

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-15T18:20:45.250388Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T18:20:45.250388Z digest=sha256:4dfe663cbb4de9f34096b5b488992f5f3cb6c16910a4b53cb9784db9fa1df9d9

Observation 1c593a86-beff-412b-ae42-05b63875bdfb · outbound

This paper cites The complexity of M arkov D ecision P rocesses.

Remembering the Markov Property in Cooperative MARL The complexity of M arkov D ecision P rocesses

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T18:20:46.046766Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-15T18:20:45.261928Z digest=sha256:cf55c458822e5f310d8e15fddc5785aa984664e8bdd97ce66775ee74be2a97da

Observation 4dab9ba7-2994-4682-9406-65442990f983 · outbound

This paper cites Benchmarking Multi-Agent Deep Reinforcement Learning Algorithms in Cooperative Tasks.

Remembering the Markov Property in Cooperative MARL Benchmarking Multi-Agent Deep Reinforcement Learning Algorithms in Cooperative Tasks

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-15T18:20:45.269096Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T18:20:45.269096Z digest=sha256:33992293eb202051ddd1856b645b37567a23661c7571c0efc76ac6fcdf1e0d9b

Observation d19a400a-4fe7-4fe0-b2f5-e0dda961b21d · outbound

This paper cites Agent modelling under partial observability for deep reinforcement learning.

Remembering the Markov Property in Cooperative MARL Agent modelling under partial observability for deep reinforcement learning

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T18:20:46.006244Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-15T18:20:45.275751Z digest=sha256:d96b3a54bb002dffc1849f9e9033160ea005e8c38d0a424fc28b4404581a93dd

Observation 2233985a-2caa-4b17-aec3-7f39164ad8bc · outbound

This paper cites Facmac: Factored multi-agent centralised policy gradients.

Remembering the Markov Property in Cooperative MARL Facmac: Factored multi-agent centralised policy gradients

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-15T18:20:45.282696Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T18:20:45.282696Z digest=sha256:524b0095e5afaa59e5653fc5e9e5ee93a7957b4945125179398ea541977c0fb1

Observation 22717ee6-b952-4100-ae0b-81ab47e12a2e · outbound

This paper cites Machine theory of mind.

Remembering the Markov Property in Cooperative MARL Machine theory of mind

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T18:20:45.955186Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-15T18:20:45.289933Z digest=sha256:d06cab950ddca76fba9d24e01eba66a3ae970378558d6736a91255d80bd7e402

Observation a98797fd-3d9a-44ff-8d22-6ee717188181 · outbound

This paper cites Mutual information between discrete and continuous data sets.

Remembering the Markov Property in Cooperative MARL Mutual information between discrete and continuous data sets

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-15T18:20:45.307175Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T18:20:45.307175Z digest=sha256:5f0866accddf4e33c15eb5b3bbfee8ec0b8fb1d0c47ccae8b120bee18a790812

Observation 0a8b6e83-a659-49fe-9f1c-cd0166d76ff9 · outbound

This paper cites JaxMARL: Multi-Agent RL Environments and Algorithms in JAX.

Remembering the Markov Property in Cooperative MARL JaxMARL: Multi-Agent RL Environments and Algorithms in JAX

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-15T18:20:45.321654Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T18:20:45.321654Z digest=sha256:610682c9e92a0fc268fc1e8f1ad58775776a1b7fcd9c0c4855c53195409533fd

Observation 45ed4b3f-9170-4581-969c-a16ff0bef7a7 · outbound

This paper cites The StarCraft Multi-Agent Challenge.

Remembering the Markov Property in Cooperative MARL The StarCraft Multi-Agent Challenge

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-15T18:20:45.335420Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T18:20:45.335420Z digest=sha256:7c7db89d3c3740cccf9d48691c123e502d49f794e39cd83a8a945cbfccc9f7e1

Observation 4824165f-4a2a-4b82-bedc-ec848fa0da91 · outbound

This paper cites Sutton and Andrew G.

Remembering the Markov Property in Cooperative MARL Sutton and Andrew G

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-15T18:20:45.344743Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T18:20:45.344743Z digest=sha256:2dfbea61a4da4e6cb31959ac70260d258c2d3c939ee93d28ade607b09802a45e

Observation ce931b36-4ec8-4cbe-8a96-ee0ae306df07 · outbound

This paper cites Order matters: Agent-by-agent policy optimization.

Remembering the Markov Property in Cooperative MARL Order matters: Agent-by-agent policy optimization

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T18:20:45.853665Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-15T18:20:45.358732Z digest=sha256:51a363a57d7dbe380b4b9d3c6a5f5ba0dbe38524c1415750492a9fdb672280d4

Observation 04b3004c-d772-4bcc-a29e-1e54b140abac · outbound

This paper cites Emergence of maps in the memories of blind navigation agents.

Remembering the Markov Property in Cooperative MARL Emergence of maps in the memories of blind navigation agents

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T18:20:45.834884Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-15T18:20:45.371678Z digest=sha256:21ed194a3fee903291ab1a991b5bfe6ac4070e53a30334e41486c3eefa8c03f9

Observation be67f0dc-42f3-489e-86ca-122c136aa495 · outbound

This paper cites Learning latent representations to influence multi-agent interaction.

Remembering the Markov Property in Cooperative MARL Learning latent representations to influence multi-agent interaction

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T18:20:45.812784Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-15T18:20:45.377556Z digest=sha256:be2ff1a8427b73a08a87aceb7d870579e827e04469858daa360683a412addbca

Observation 95e00ed8-4c54-4fbc-bc6c-e5f12e2e37ea · outbound

This paper cites The surprising effectiveness of ppo in cooperative multi-agent games.

Remembering the Markov Property in Cooperative MARL The surprising effectiveness of ppo in cooperative multi-agent games

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-15T18:20:45.387404Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T18:20:45.387404Z digest=sha256:28da80943d60d22b790176d666de2f537019eac6f643b6e0557af574c4daa3c8

Observation 8d1b8beb-e096-4cce-947e-48a0ebd00ea3 · outbound

This paper cites Heterogeneous-agent reinforcement learning.

Remembering the Markov Property in Cooperative MARL Heterogeneous-agent reinforcement learning

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-15T18:20:45.393733Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T18:20:45.393733Z digest=sha256:6332be2f75e2865c37cb6f6620ccbebe503a2d503f3e1d831d7ee97410d57400

Observation da94f39f-e44c-43f5-9bee-21cc1147454d · outbound

This paper cites Deep interactive bayesian reinforcement learning via meta-learning.

Remembering the Markov Property in Cooperative MARL Deep interactive bayesian reinforcement learning via meta-learning

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T18:20:45.757902Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-15T18:20:45.408594Z digest=sha256:7f3afac6ff64cba3af3306538e485d25e260032aeef119dd923b51d61c37a02c

Observation 13b01a95-80e2-4ca4-9965-6cdd89b796e4 · outbound

This paper cites write newline.

Remembering the Markov Property in Cooperative MARL write newline

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-15T18:20:45.426207Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T18:20:45.426207Z digest=sha256:a658e9a68f2616be69141a3e2a19d6076aa192ea998fa01ab5d9fa09944e0e94

Pith citing papers

No inbound Pith citation observations are available.