Pith. sign in

Paper Citation Record · LEDGER

Diffusing to Coordinate: Efficient Online Multi-Agent Diffusion Policies

As of 14 August 2026, this Paper Citation Record lists 78 of 78 outbound references and 1 inbound Pith citation observation for arXiv:2602.18291.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2602.18291 v2

Coverage vector

measured 78 of 78 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-02T22:04:28.337206Z

measured 79 of 79 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-14T06:32:32.682623+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-05-13T02:09:08.378144Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-13T02:17:07.063482Z

Reference resolution

78 of 78 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved78
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 8aa0585d-5ab1-4f53-8af0-466412b8de62 · outbound

This paper cites Reducing over- estimation bias in multi-agent domains using double centralized critics.Advances in Neural Information Processing Systems: Deep Reinforcement Learning Workshop, 2019.

Diffusing to Coordinate: Efficient Online Multi-Agent Diffusion Policies Reducing over- estimation bias in multi-agent domains using double centralized critics.Advances in Neural Information Processing Systems: Deep Reinforcement Learning Workshop, 2019

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-02T22:04:21.353437Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T22:04:21.353437Z digest=sha256:cb4b9a485af73c2053bda15a5a5854e85d2ac84f0e9b09599cb64c5bf7cef2a1

Observation 0edcfc78-4495-4699-82d9-b2bec4ede4b5 · outbound

This paper cites Is conditional generative modeling all you need for decision making? InThe Eleventh International Conference on Learning Representations, 2023.

Diffusing to Coordinate: Efficient Online Multi-Agent Diffusion Policies Is conditional generative modeling all you need for decision making? InThe Eleventh International Conference on Learning Representations, 2023

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-02T22:04:21.409314Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T22:04:21.409314Z digest=sha256:a02b75d9d624f87dcb4e09f84ab404fd72115ea58556755ef8f9602435b3e0ec

Observation d463945b-6886-47fa-a9b5-5b6c964e2a34 · outbound

This paper cites Heterogeneous agent q-weighted policy optimization.

Diffusing to Coordinate: Efficient Online Multi-Agent Diffusion Policies Heterogeneous agent q-weighted policy optimization

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-02T22:04:21.554562Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T22:04:21.554562Z digest=sha256:4812e78c86cf90cd186773ad56b1d70da66760ee18e39e9e71e017590baa9d68

Observation 09a54f33-e86c-405e-a452-c8e610f0edd2 · outbound

This paper cites A distributional perspective on reinforce- ment learning.

Diffusing to Coordinate: Efficient Online Multi-Agent Diffusion Policies A distributional perspective on reinforce- ment learning

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-02T22:04:21.658257Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T22:04:21.658257Z digest=sha256:fa7ae8f366f3af6acbc6c659f9b7c3c678dfaa361b0c35a84615b51e3323b178

Observation 07baafda-2862-491e-83a4-c8976017d7f5 · outbound

This paper cites The complexity of decentralized control of markov decision processes.Mathematics of operations research, 27(4):819–840, 2002.

Diffusing to Coordinate: Efficient Online Multi-Agent Diffusion Policies The complexity of decentralized control of markov decision processes.Mathematics of operations research, 27(4):819–840, 2002

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-02T22:04:21.790972Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T22:04:21.790972Z digest=sha256:ba02bd5762e616e43e960f584a9704144c7722cec4e85e7ecc3c18cae79b9ad6

Observation 5e76e748-7506-4273-8dbb-6bf2e7cfa2f3 · outbound

This paper cites Crossq: Batch normalization in deep reinforcement learning for greater sample efficiency and simplicity.

Diffusing to Coordinate: Efficient Online Multi-Agent Diffusion Policies Crossq: Batch normalization in deep reinforcement learning for greater sample efficiency and simplicity

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-02T22:04:21.886299Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T22:04:21.886299Z digest=sha256:25bf9c96f702feeb6a5b2988d8168f264bfb24c6f2fd31e0112d11354e39f68c

Observation fda59a7a-484e-4ba2-9acf-fe379e1fc0b1 · outbound

This paper cites Video generation models as world simulators.OpenAI Blog, 2024.

Diffusing to Coordinate: Efficient Online Multi-Agent Diffusion Policies Video generation models as world simulators.OpenAI Blog, 2024

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-02T22:04:22.024329Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T22:04:22.024329Z digest=sha256:345707e1616ed166b3a7459e4cd36396c0e6cf7580c3052005e766a8a8df339a

Observation 67c50d4b-b926-4457-a400-8f085394252d · outbound

This paper cites Dime: Diffusion-based maximum entropy reinforcement learn- ing.

Diffusing to Coordinate: Efficient Online Multi-Agent Diffusion Policies Dime: Diffusion-based maximum entropy reinforcement learn- ing

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-02T22:04:22.107650Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T22:04:22.107650Z digest=sha256:de6f29cf61c5e03a06365052b37331f1924ce330df2fa82e35a06589f31eb9d8

Observation 4632eb1b-4764-4a87-be1c-d9fca1a47a7c · outbound

This paper cites Offline reinforcement learn- ing via high-fidelity generative behavior modeling.

Diffusing to Coordinate: Efficient Online Multi-Agent Diffusion Policies Offline reinforcement learn- ing via high-fidelity generative behavior modeling

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-02T22:04:22.248648Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T22:04:22.248648Z digest=sha256:2d1947b5dc42e05d6f2e2ec164029a18cbcfb441d9500e93e2c239b702cc836e

Observation 799e4683-82bc-47e2-b4e7-6b1fa48ebdf6 · outbound

This paper cites Multi-agent systems for robotic autonomy with llms.

Diffusing to Coordinate: Efficient Online Multi-Agent Diffusion Policies Multi-agent systems for robotic autonomy with llms

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-02T22:04:22.303616Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T22:04:22.303616Z digest=sha256:048fe056cb67fc88828b228d2eca746dfbff42e7e513299a75fb83d43dd448d6

Observation d61f4dbc-23d7-4bfc-baef-afb2bd27b1cf · outbound

This paper cites Novelty-guided data reuse for efficient and diversified multi-agent reinforcement learning.

Diffusing to Coordinate: Efficient Online Multi-Agent Diffusion Policies Novelty-guided data reuse for efficient and diversified multi-agent reinforcement learning

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-02T22:04:22.362502Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T22:04:22.362502Z digest=sha256:296a18f57a68126d8fc6bd5a9220a6996ee7a89dad014fd38f46150756287e2f

Observation 1b96bde8-613c-410c-9406-c529ea269af3 · outbound

This paper cites Continuous q-score matching: Diffusion guided reinforcement learning for continuous-time control.

Diffusing to Coordinate: Efficient Online Multi-Agent Diffusion Policies Continuous q-score matching: Diffusion guided reinforcement learning for continuous-time control

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-02T22:04:22.420165Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T22:04:22.420165Z digest=sha256:769376a3473ac4ccf6c38d470837f58caa7bcd16e192950c99d1988b38add356

Observation 82761d8e-209e-4384-ab04-0a31f903b255 · outbound

This paper cites Diffusion policy: Visuomotor policy learning via action diffusion.

Diffusing to Coordinate: Efficient Online Multi-Agent Diffusion Policies Diffusion policy: Visuomotor policy learning via action diffusion

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-02T22:04:22.495890Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T22:04:22.495890Z digest=sha256:ca9527ae4b0474e018bf0c35b48c1274e3b3605eecb1a48f664eade9c8ddb804

Observation b05aa95b-e6cf-4be0-823d-dde141e83ccf · outbound

This paper cites Is Independent Learning All You Need in the StarCraft Multi-Agent Challenge?.

Diffusing to Coordinate: Efficient Online Multi-Agent Diffusion Policies Is Independent Learning All You Need in the StarCraft Multi-Agent Challenge?

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-02T22:04:22.549456Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T22:04:22.549456Z digest=sha256:126dd1d481d561041daa3aea15ca0bc56b7fe153f8c29f40258b6d0dc380f91a

Observation b3af9bd1-ab6a-4233-8f63-7eeb5f421f7d · outbound

This paper cites Diffusion-based reinforcement learning via q-weighted variational policy optimization.

Diffusing to Coordinate: Efficient Online Multi-Agent Diffusion Policies Diffusion-based reinforcement learning via q-weighted variational policy optimization

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-02T22:04:22.616890Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T22:04:22.616890Z digest=sha256:f285bf32c2b9939da7e676e392f999461fea50e586874c04f783a26f5cd894ee

Observation 663ce302-8b4f-484e-9dd1-fb29515736c6 · outbound

This paper cites Maximum entropy reinforcement learning with diffusion policy.

Diffusing to Coordinate: Efficient Online Multi-Agent Diffusion Policies Maximum entropy reinforcement learning with diffusion policy

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-02T22:04:22.661495Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T22:04:22.661495Z digest=sha256:55a41b8e59f3ad03d26f76f9b98d3919247a431b08c8b4ed72e1c5599ed1e874

Observation 0bd1f0fe-d024-49bb-bdb5-0740687bac60 · outbound

This paper cites Diffusion actor-critic: Formulating constrained policy iteration as diffusion noise regression for offline reinforcement learning.

Diffusing to Coordinate: Efficient Online Multi-Agent Diffusion Policies Diffusion actor-critic: Formulating constrained policy iteration as diffusion noise regression for offline reinforcement learning

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-02T22:04:22.736830Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T22:04:22.736830Z digest=sha256:3b65a9a88b63a792be48b9c22b2c514a9d5c087b6873d9b63fa4a44d6637fa0e

Observation 0f49e7df-6b59-42e2-b179-a16c2dc9b560 · outbound

This paper cites Counterfactual multi-agent policy gradients.

Diffusing to Coordinate: Efficient Online Multi-Agent Diffusion Policies Counterfactual multi-agent policy gradients

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-02T22:04:22.792076Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T22:04:22.792076Z digest=sha256:9b5864a455899b2bfb389c91e719983e99244ff6bba15163d055dafe90ac207a

Observation 3bd49bc0-0bbe-473b-8f4a-7e2950fe5051 · outbound

This paper cites INS: Interaction-aware synthesis to enhance offline multi-agent reinforcement learning.

Diffusing to Coordinate: Efficient Online Multi-Agent Diffusion Policies INS: Interaction-aware synthesis to enhance offline multi-agent reinforcement learning

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-02T22:04:22.868522Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T22:04:22.868522Z digest=sha256:4b06d0020b9dfe9f42ec5f35644aa34bd706fc2db7873ebe97d74a9268e6e52d

Observation ac1dc2d7-cc3d-4f59-b3ff-3af4bf9df954 · outbound

This paper cites Reinforcement learning with deep energy-based policies.

Diffusing to Coordinate: Efficient Online Multi-Agent Diffusion Policies Reinforcement learning with deep energy-based policies

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-02T22:04:22.972612Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T22:04:22.972612Z digest=sha256:89cdff3859bdfb67bed6db38a174a73583eabec52fe35d21af9d8a1eccbd1b8d

Observation fbeb0abf-df4c-4d79-a5f3-8d74e0f4cf14 · outbound

This paper cites Soft actor-critic: Off- policy maximum entropy deep reinforcement learning with a stochastic actor.

Diffusing to Coordinate: Efficient Online Multi-Agent Diffusion Policies Soft actor-critic: Off- policy maximum entropy deep reinforcement learning with a stochastic actor

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-02T22:04:23.058320Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T22:04:23.058320Z digest=sha256:9e60e0133227fd2f8ecba2dbf9c8ebd6eb748f8e8d818f662097cb525549a498

Observation 2e9b247a-b8c4-4db6-b760-e86f77e3b138 · outbound

This paper cites IDQL: Implicit Q-Learning as an Actor-Critic Method with Diffusion Policies.

Diffusing to Coordinate: Efficient Online Multi-Agent Diffusion Policies IDQL: Implicit Q-Learning as an Actor-Critic Method with Diffusion Policies

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-02T22:04:23.206034Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T22:04:23.206034Z digest=sha256:aa9372357080b3ba582705536d35b8b84ba4755c12ccc334fc8fb7b689e72ca8

Observation 85723d21-12ca-48d9-96ba-a0e76320a84b · outbound

This paper cites Denoising diffusion probabilistic models.Advances in Neural Information Processing Systems, 33:6840–6851, 2020.

Diffusing to Coordinate: Efficient Online Multi-Agent Diffusion Policies Denoising diffusion probabilistic models.Advances in Neural Information Processing Systems, 33:6840–6851, 2020

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-02T22:04:23.370859Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T22:04:23.370859Z digest=sha256:c5895a4be9e3a7b7b535a5a01e0928f257f35c794c7a6d84857c3bc128ed4bcf

Observation 5bbb4517-1014-499d-8767-4994dfd49b5a · outbound

This paper cites Value diffusion reinforcement learning.

Diffusing to Coordinate: Efficient Online Multi-Agent Diffusion Policies Value diffusion reinforcement learning

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-02T22:04:23.446965Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T22:04:23.446965Z digest=sha256:cd8b2b098da28358563722e160236914df19b98878802d226b9d62595ad10738

Observation 43490b08-84e7-47af-adb1-e147b1267be1 · outbound

This paper cites Diffuseloco: Real-time legged locomotion control with diffusion from offline datasets.

Diffusing to Coordinate: Efficient Online Multi-Agent Diffusion Policies Diffuseloco: Real-time legged locomotion control with diffusion from offline datasets

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-02T22:04:23.540653Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T22:04:23.540653Z digest=sha256:5143a27ad2f4b10bebede3eecf7ca230395f37b839c9c4dbcc9804c557a62a6b

Observation cb96d947-19fe-4f2c-899a-5e252f95472d · outbound

This paper cites Diffuseloco: Real-time legged locomotion control with diffusion from offline datasets.

Diffusing to Coordinate: Efficient Online Multi-Agent Diffusion Policies Diffuseloco: Real-time legged locomotion control with diffusion from offline datasets

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-02T22:04:23.580921Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T22:04:23.580921Z digest=sha256:aed8b5c31194ce6dbaecd004846bc4e761757829efe6b1d90e02bf454c1e5980

Observation feb40cb8-1c11-4ec3-987f-487a6281f3f3 · outbound

This paper cites Planning with diffusion for flexible behavior synthesis.

Diffusing to Coordinate: Efficient Online Multi-Agent Diffusion Policies Planning with diffusion for flexible behavior synthesis

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-02T22:04:23.626124Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T22:04:23.626124Z digest=sha256:8ed96f002122bae1c3cc658d38c35af65a312ae000cda75940f2b121d7e87bea

Observation 9ab555b2-6de4-453c-9a13-d710bd142a22 · outbound

This paper cites Agent- centric actor-critic for asynchronous multi-agent reinforcement learning.

Diffusing to Coordinate: Efficient Online Multi-Agent Diffusion Policies Agent- centric actor-critic for asynchronous multi-agent reinforcement learning

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-02T22:04:23.703971Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T22:04:23.703971Z digest=sha256:b869988738520281aabd8060b98f611c3e57f60dae2b5067310c25f28365d910

Observation 8f0679dd-387c-4ce2-b8ca-edcb5a2077a6 · outbound

This paper cites Efficient diffusion poli- cies for offline reinforcement learning.Advances in Neural Information Processing Systems, 36:67195–67212, 2023.

Diffusing to Coordinate: Efficient Online Multi-Agent Diffusion Policies Efficient diffusion poli- cies for offline reinforcement learning.Advances in Neural Information Processing Systems, 36:67195–67212, 2023

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-02T22:04:23.746367Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T22:04:23.746367Z digest=sha256:d61e388ce25c6f2364e7b50a7f1c0ee9a965dadb7779df17c69bbb39fe321b78

Observation d90a251c-2c0b-4ef2-b90a-b2fd4f02f819 · outbound

This paper cites Enhancing cooperative multi-agent reinforcement learning with state modelling and adversarial exploration.

Diffusing to Coordinate: Efficient Online Multi-Agent Diffusion Policies Enhancing cooperative multi-agent reinforcement learning with state modelling and adversarial exploration

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-02T22:04:23.793464Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T22:04:23.793464Z digest=sha256:2b547ddd6912db78ee5b190e08ac77697f83f91556516636d45907fe1a8b4e0a

Observation ffae433b-0756-43ff-af77-81ac47cfdee5 · outbound

This paper cites Gta: Generative trajectory aug- mentation with guidance for offline reinforcement learning.Advances in Neural Information Processing Systems, 37:56766–56801, 2024.

Diffusing to Coordinate: Efficient Online Multi-Agent Diffusion Policies Gta: Generative trajectory aug- mentation with guidance for offline reinforcement learning.Advances in Neural Information Processing Systems, 37:56766–56801, 2024

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-02T22:04:23.907200Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T22:04:23.907200Z digest=sha256:d94aaa6f9c0966a6d4ca4ab18707dfa6836d5ae3a30c737db089d7e35d859710

Observation b0374e21-61ad-43a9-baa3-3cb0443f0e01 · outbound

This paper cites Dof: A diffusion factorization framework for offline multi-agent decision making.

Diffusing to Coordinate: Efficient Online Multi-Agent Diffusion Policies Dof: A diffusion factorization framework for offline multi-agent decision making

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-02T22:04:23.997509Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T22:04:23.997509Z digest=sha256:74a541bfc41d1a53cf79af777cbea364ea6aab39e92bbfa2450a5e76d1d0240d

Observation e3fed53a-03fc-4507-ac70-a62003398cae · outbound

This paper cites Race: improve multi-agent reinforcement learning with representation asymmetry and collaborative evolution.

Diffusing to Coordinate: Efficient Online Multi-Agent Diffusion Policies Race: improve multi-agent reinforcement learning with representation asymmetry and collaborative evolution

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-02T22:04:24.134696Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T22:04:24.134696Z digest=sha256:80fd5d49b0d7e53e0b2a965f702844d1c2e20c96de7918e8fa220d3372a5e8ad

Observation 44cd21bf-5ce7-40c3-bccc-de5beccb5af5 · outbound

This paper cites Revisiting cooperative off-policy multi-agent reinforcement learning.

Diffusing to Coordinate: Efficient Online Multi-Agent Diffusion Policies Revisiting cooperative off-policy multi-agent reinforcement learning

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-02T22:04:24.290599Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T22:04:24.290599Z digest=sha256:5811d6ccadeb2821c5c7b27dc6ccf8c6b2d56374ff816916add7daacb30e8b78

Observation ba7aa86f-93e3-421b-ae74-7263d2ce5661 · outbound

This paper cites Beyond conservatism: Diffusion policies in offline multi-agent reinforcement learning, 2023.

Diffusing to Coordinate: Efficient Online Multi-Agent Diffusion Policies Beyond conservatism: Diffusion policies in offline multi-agent reinforcement learning, 2023

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-02T22:04:24.463731Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T22:04:24.463731Z digest=sha256:a01d773467fb71564f72b66ef46f701c54cf012771fe95a2e7b61bb4f87dae6b

Observation c1a30748-be35-465d-8249-93e81cb93cb9 · outbound

This paper cites Maximum entropy heterogeneous-agent reinforcement learning.

Diffusing to Coordinate: Efficient Online Multi-Agent Diffusion Policies Maximum entropy heterogeneous-agent reinforcement learning

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-02T22:04:24.577770Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T22:04:24.577770Z digest=sha256:d8cad548d4645bacb69a042e0ed6ee0dc6ea102614af3f5db7f0ccd47ae865ca

Observation da47b002-9793-4460-9c1a-be8802cfcd34 · outbound

This paper cites Multi-agent actor-critic for mixed cooperative-competitive environments.Advances in neural information processing systems, 30, 2017.

Diffusing to Coordinate: Efficient Online Multi-Agent Diffusion Policies Multi-agent actor-critic for mixed cooperative-competitive environments.Advances in neural information processing systems, 30, 2017

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-02T22:04:24.720043Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T22:04:24.720043Z digest=sha256:5af3ea40f8f3f0c0b66edd600b22897c97fd9d940c9f70ab5c88d64ae8e1e453

Observation adcda0d4-dfeb-4616-99e1-f4f6b480de7c · outbound

This paper cites Contrastive energy prediction for exact energy-guided diffusion sampling in offline reinforcement learning.

Diffusing to Coordinate: Efficient Online Multi-Agent Diffusion Policies Contrastive energy prediction for exact energy-guided diffusion sampling in offline reinforcement learning

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-02T22:04:24.825764Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T22:04:24.825764Z digest=sha256:9363112c00d103b629d1d8072d12f6ce364e1b798a520c0ddf2b4ac49180cd54

Observation af3f3405-2da2-4785-a63c-7529943d029e · outbound

This paper cites Dpm-solver: A fast ode solver for diffusion probabilistic model sampling in around 10 steps.Advances in neural information processing systems, 35:5775–5787, 2022.

Diffusing to Coordinate: Efficient Online Multi-Agent Diffusion Policies Dpm-solver: A fast ode solver for diffusion probabilistic model sampling in around 10 steps.Advances in neural information processing systems, 35:5775–5787, 2022

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-02T22:04:24.986702Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T22:04:24.986702Z digest=sha256:5ff711a9e58da5091c26f9f32cd0b3e185f100d348d1409f0bdad8fbdf266dc4

Observation 428fa515-8be7-42a9-b45d-3d8b48460055 · outbound

This paper cites Synthetic experience replay.

Diffusing to Coordinate: Efficient Online Multi-Agent Diffusion Policies Synthetic experience replay

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-02T22:04:25.087793Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T22:04:25.087793Z digest=sha256:ef3f54070344e8fa8eb0d74eb897233c2328f02c0c2c3e457d0eba79c24f158f

Observation 01b9b2a9-f57b-4b5a-b520-741750ccfe74 · outbound

This paper cites Efficient online reinforcement learning for diffusion policy.

Diffusing to Coordinate: Efficient Online Multi-Agent Diffusion Policies Efficient online reinforcement learning for diffusion policy

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-02T22:04:25.240463Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T22:04:25.240463Z digest=sha256:69d70f3d5774d6bf5e092d66738e2f954952320e967e36ba6f0cda02e533de3e

Observation 32954f35-5902-4992-92b3-a1183ba4017a · outbound

This paper cites Coordinated multi-robot exploration under communication constraints using decentralized markov decision processes.

Diffusing to Coordinate: Efficient Online Multi-Agent Diffusion Policies Coordinated multi-robot exploration under communication constraints using decentralized markov decision processes

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-02T22:04:25.376586Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T22:04:25.376586Z digest=sha256:4fb52e79af4f4691ff9c004dfdd92fc572d3245033dd2797880a8fc8febd432c

Observation 0db3259b-bd14-4f20-8970-3a20b242a678 · outbound

This paper cites Springer, 2016.

Diffusing to Coordinate: Efficient Online Multi-Agent Diffusion Policies Springer, 2016

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-02T22:04:25.507543Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T22:04:25.507543Z digest=sha256:e873c3cf1a74ba0d8ff44c0851f42019c78d1389690cbcb53c144833a793885a

Observation ac5dc06c-64e2-4f20-80de-c15f7ca36f90 · outbound

This paper cites Optimal and approximate q-value functions for decentralized pomdps.Journal of Artificial Intelligence Research, 32:289–353, 2008.

Diffusing to Coordinate: Efficient Online Multi-Agent Diffusion Policies Optimal and approximate q-value functions for decentralized pomdps.Journal of Artificial Intelligence Research, 32:289–353, 2008

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-02T22:04:25.654360Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T22:04:25.654360Z digest=sha256:826bf4952bcce727f1fea2638f2049ab9bcb7e5ec953795ae213a6517fa7c958

Observation b0332c0e-26e0-49f1-bcfc-09de53732ba9 · outbound

This paper cites Regularized softmax deep multi-agent q-learning.Advances in Neural Information Processing Systems, 34:1365–1377, 2021.

Diffusing to Coordinate: Efficient Online Multi-Agent Diffusion Policies Regularized softmax deep multi-agent q-learning.Advances in Neural Information Processing Systems, 34:1365–1377, 2021

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-02T22:04:25.787566Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T22:04:25.787566Z digest=sha256:d1d765b1974ae22d99bd3b0cbaaca197469db50147f205972a5cd7961d24e1d3

Observation 006b0cce-ab73-406f-87c5-d3600ac8bc26 · outbound

This paper cites Imitating human behaviour with diffusion models.

Diffusing to Coordinate: Efficient Online Multi-Agent Diffusion Policies Imitating human behaviour with diffusion models

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-02T22:04:25.923902Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T22:04:25.923902Z digest=sha256:f49d83d12114dd0bc1f00ed51dc784908c6ec0b2c97366276879c7e00a549601

Observation a49f2ac6-cd40-46ab-a13b-72fc3f204ef6 · outbound

This paper cites Facmac: Factored multi-agent centralised policy gradients.Advances in Neural Information Processing Systems, 34:12208–12221, 2021.

Diffusing to Coordinate: Efficient Online Multi-Agent Diffusion Policies Facmac: Factored multi-agent centralised policy gradients.Advances in Neural Information Processing Systems, 34:12208–12221, 2021

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-02T22:04:26.069456Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T22:04:26.069456Z digest=sha256:362516b2bc4c57bdd68b0988f8cb75dcc2c4e373f3b09e8235b7fffd7558d1fd

Observation b33d97f5-84e8-4b4f-bb7a-7b0cf6f1551c · outbound

This paper cites Learning a diffusion model policy from rewards via q-score matching.

Diffusing to Coordinate: Efficient Online Multi-Agent Diffusion Policies Learning a diffusion model policy from rewards via q-score matching

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-02T22:04:26.187590Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T22:04:26.187590Z digest=sha256:bbbef990d4994c051519e0447f06558e5d5f5cee0922782cea43fcf668864fd6

Observation c615cfca-10b0-424a-8515-20246f653564 · outbound

This paper cites Offline multi- agent reinforcement learning via score decomposition, 2025.

Diffusing to Coordinate: Efficient Online Multi-Agent Diffusion Policies Offline multi- agent reinforcement learning via score decomposition, 2025

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-02T22:04:26.258377Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T22:04:26.258377Z digest=sha256:07ad078e399e94f530626b9e4205736de3fc6f0e99d2932ccef217051423e8a2

Observation 8354513a-2915-49a5-87bc-5397a5c11fdc · outbound

This paper cites Qmix: Monotonic value function factorisation for deep multi-agent reinforcement learning.

Diffusing to Coordinate: Efficient Online Multi-Agent Diffusion Policies Qmix: Monotonic value function factorisation for deep multi-agent reinforcement learning

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-02T22:04:26.299270Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T22:04:26.299270Z digest=sha256:c4fa75ec0e5e8e79d786d219f2c137f6a3b9c3e391f63e317b83fd877eb0833b

Observation 1dfe89fe-c934-43eb-815a-1ad402d9021c · outbound

This paper cites Diffusion policy policy op- timization.

Diffusing to Coordinate: Efficient Online Multi-Agent Diffusion Policies Diffusion policy policy op- timization

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-02T22:04:26.376396Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T22:04:26.376396Z digest=sha256:9a9e8014ec9d64447cd554f64344fb5f144255a08f99d0bf2d68c9cc0d8a2570

Observation 18573499-073d-4276-abcf-a5d0827711ba · outbound

This paper cites Cambridge University Press, 2019.

Diffusing to Coordinate: Efficient Online Multi-Agent Diffusion Policies Cambridge University Press, 2019

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-02T22:04:26.433524Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T22:04:26.433524Z digest=sha256:a4b681c499fcf5ca80a1a7b502ee1b7cdd0ff178e5f762ca06fb05c3cae5090e

Observation 09732023-1356-4ad4-b559-4124bd17bb6f · outbound

This paper cites Deep unsuper- vised learning using nonequilibrium thermodynamics.

Diffusing to Coordinate: Efficient Online Multi-Agent Diffusion Policies Deep unsuper- vised learning using nonequilibrium thermodynamics

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-02T22:04:26.477404Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T22:04:26.477404Z digest=sha256:b0694f2c3fcae32250dcd6f633cc528982b9e730c0075882b7b2004e1a4b39c5

Observation a5d05d26-4c9a-4bd1-887c-d167ef1cca15 · outbound

This paper cites Qtran: Learning to factorize with transformation for cooperative multi-agent reinforcement learning.

Diffusing to Coordinate: Efficient Online Multi-Agent Diffusion Policies Qtran: Learning to factorize with transformation for cooperative multi-agent reinforcement learning

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-02T22:04:26.523923Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T22:04:26.523923Z digest=sha256:c7225a78619cd38fc688e94879005ed62f87afe8ff0747aabd3464d30bb164c3

Observation fc62605d-a25b-4575-bd5a-48920336f2f8 · outbound

This paper cites Generative modeling by estimating gradients of the data distribution.Advances in Neural Information Processing Systems, 32, 2019.

Diffusing to Coordinate: Efficient Online Multi-Agent Diffusion Policies Generative modeling by estimating gradients of the data distribution.Advances in Neural Information Processing Systems, 32, 2019

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-02T22:04:26.560397Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T22:04:26.560397Z digest=sha256:5597d0db3af1b766598b3ab959a3f0704d6bfbc92fefdb993d2dcdf6a54ea2da

Observation 5eca49e0-f487-41b4-9a84-4b4e1ce2d1dc · outbound

This paper cites Score-based generative modeling through stochastic differential equations.

Diffusing to Coordinate: Efficient Online Multi-Agent Diffusion Policies Score-based generative modeling through stochastic differential equations

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-02T22:04:26.636123Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T22:04:26.636123Z digest=sha256:f616940b3e5e4d430b7974c56386635e9c78bbfc520a4d3c5b947b8de5c051ed

Observation 22a449f3-5bee-405b-8a21-f0d3095deb35 · outbound

This paper cites Value- decomposition networks for cooperative multi-agent learning based on team reward.

Diffusing to Coordinate: Efficient Online Multi-Agent Diffusion Policies Value- decomposition networks for cooperative multi-agent learning based on team reward

Reference 57

Resolution
unresolved
no resolver link, observed 2026-08-02T22:04:26.722252Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T22:04:26.722252Z digest=sha256:b1c5a034361c28cd641aa124366dcbaf1f2819558ab2e9e4653d9612dfa77e53

Observation 69d0f040-3a44-49c5-a2f4-0e19ca4ccbc9 · outbound

This paper cites Multiagent cooperation and competition with deep reinforcement learning.PloS one, 12(4):e0172395, 2017.

Diffusing to Coordinate: Efficient Online Multi-Agent Diffusion Policies Multiagent cooperation and competition with deep reinforcement learning.PloS one, 12(4):e0172395, 2017

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-02T22:04:26.825836Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T22:04:26.825836Z digest=sha256:ee865a9d1e640224678f60cc0d5634722f4a2733614551caec35615d6d758918

Observation 5e072159-5d89-4d28-be9f-8cfd2f07a4ba · outbound

This paper cites Multi-agent reinforcement learning: Independent vs.

Diffusing to Coordinate: Efficient Online Multi-Agent Diffusion Policies Multi-agent reinforcement learning: Independent vs

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-02T22:04:26.926284Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T22:04:26.926284Z digest=sha256:034f56e82e4d11f2898b29c741281d92bf28ecafbbb6557dbf5c606fe318a37e

Observation 05cf5ec7-49c5-4160-af66-449e38478cc2 · outbound

This paper cites Qplex: Duplex dueling multi-agent q-learning.

Diffusing to Coordinate: Efficient Online Multi-Agent Diffusion Policies Qplex: Duplex dueling multi-agent q-learning

Reference 60

Resolution
unresolved
no resolver link, observed 2026-08-02T22:04:27.018340Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T22:04:27.018340Z digest=sha256:d30c9f277ed326662c8fcd028bb2a2a66a07c276f8addb19664dade062580445

Observation 5ef2290f-b7b6-4e14-8ac3-b95d98e4fdd8 · outbound

This paper cites Diffusion actor-critic with entropy regulator.

Diffusing to Coordinate: Efficient Online Multi-Agent Diffusion Policies Diffusion actor-critic with entropy regulator

Reference 61

Resolution
unresolved
no resolver link, observed 2026-08-02T22:04:27.084343Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T22:04:27.084343Z digest=sha256:13e7db2f8c370280b878a4112a4cc22e1b6b937d7912f313a03f220d77898848

Observation 8c3d0d84-a22a-4091-b3c1-a46bdf9e3f08 · outbound

This paper cites Enhanced dacer algorithm with high diffusion efficiency.arXiv preprint arXiv:2505.23426, 2025.

Diffusing to Coordinate: Efficient Online Multi-Agent Diffusion Policies Enhanced dacer algorithm with high diffusion efficiency.arXiv preprint arXiv:2505.23426, 2025

Reference 62

Resolution
unresolved
no resolver link, observed 2026-08-02T22:04:27.162734Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T22:04:27.162734Z digest=sha256:a960e9da70c88dbcf845afdba23380732af5b3b6ab8ef93f6f1e231b833d249a

Observation a42577d2-d03c-4c23-aa9d-686f5adf05fe · outbound

This paper cites Diffusion policies as an expressive policy class for offline reinforcement learning.

Diffusing to Coordinate: Efficient Online Multi-Agent Diffusion Policies Diffusion policies as an expressive policy class for offline reinforcement learning

Reference 63

Resolution
unresolved
no resolver link, observed 2026-08-02T22:04:27.227314Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T22:04:27.227314Z digest=sha256:13619731cc7e6b3b0cfd709c52209b51bea3cca1efdbc0afe4831c3b31fb7848

Observation 22d90049-e7b9-483d-a551-58697a16d129 · outbound

This paper cites Learning intractable multimodal policies with repa- rameterization and diversity regularization.

Diffusing to Coordinate: Efficient Online Multi-Agent Diffusion Policies Learning intractable multimodal policies with repa- rameterization and diversity regularization

Reference 64

Resolution
unresolved
no resolver link, observed 2026-08-02T22:04:27.304994Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T22:04:27.304994Z digest=sha256:2bfec49d17794279110eab0811e375522048f883a206114e36c31cc7c409bc84

Observation 75fbc84b-3ef7-40ad-969f-1f355fa74897 · outbound

This paper cites Latent diffusion planning for imitation learning.

Diffusing to Coordinate: Efficient Online Multi-Agent Diffusion Policies Latent diffusion planning for imitation learning

Reference 65

Resolution
unresolved
no resolver link, observed 2026-08-02T22:04:27.385235Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T22:04:27.385235Z digest=sha256:7e3aeae2a3e148f280da3514c67a5419c41f45dcd7cac438639ab050bbbf04af

Observation a573defa-0370-4ce4-832b-3a1ef0e00e5a · outbound

This paper cites Multi-agent reinforcement learning with communication-constrained priors.

Diffusing to Coordinate: Efficient Online Multi-Agent Diffusion Policies Multi-agent reinforcement learning with communication-constrained priors

Reference 66

Resolution
unresolved
no resolver link, observed 2026-08-02T22:04:27.452565Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T22:04:27.452565Z digest=sha256:75ce0410760ee0016486ec27b37bc29e816a8bf54be2ab63408d32f82b3d8048

Observation f6ae37c0-7199-4409-a2f7-82b83cdbeeb0 · outbound

This paper cites Policy Representation via Diffusion Probability Model for Reinforcement Learning.

Diffusing to Coordinate: Efficient Online Multi-Agent Diffusion Policies Policy Representation via Diffusion Probability Model for Reinforcement Learning

Reference 67

Resolution
unresolved
no resolver link, observed 2026-08-02T22:04:27.521008Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T22:04:27.521008Z digest=sha256:be3a16865bf9831012dd429080902073ebaf75b42158643f9a01bac5c748078f

Observation 7281ef6b-bfc0-4c19-9ccd-d63b7abe0f5c · outbound

This paper cites Fine-tuning diffusion policies with backpropagation through diffusion timesteps, 2025.

Diffusing to Coordinate: Efficient Online Multi-Agent Diffusion Policies Fine-tuning diffusion policies with backpropagation through diffusion timesteps, 2025

Reference 68

Resolution
unresolved
no resolver link, observed 2026-08-02T22:04:27.597916Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T22:04:27.597916Z digest=sha256:25ac94ac451eb4b352f255e38432e417441e7f7a54be836a935ba3e1a3e97798

Observation 38e11d57-8b9d-4527-b36b-13f48431defa · outbound

This paper cites The surprising effectiveness of ppo in cooperative multi-agent games.Advances in neural in- formation processing systems, 35:24611–24624, 2022.

Diffusing to Coordinate: Efficient Online Multi-Agent Diffusion Policies The surprising effectiveness of ppo in cooperative multi-agent games.Advances in neural in- formation processing systems, 35:24611–24624, 2022

Reference 69

Resolution
unresolved
no resolver link, observed 2026-08-02T22:04:27.681681Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T22:04:27.681681Z digest=sha256:ce3351f8d8bcbc327cdbb889f9f3e5420d06ba4217bb9a1c1b1375ca64ae0b60

Observation 4aa68b3c-890f-4b5c-a5c1-255c4d0e910b · outbound

This paper cites Turbodiffusion: Accelerating video diffusion models by 100-200 times.

Diffusing to Coordinate: Efficient Online Multi-Agent Diffusion Policies Turbodiffusion: Accelerating video diffusion models by 100-200 times

Reference 70

Resolution
unresolved
no resolver link, observed 2026-08-02T22:04:27.751891Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T22:04:27.751891Z digest=sha256:1a500902ac93915fa1bbb23decde3d4d04e6d73eeacbde22ed4e7d6231b1f3ed

Observation 2eaa20ea-6bc4-4662-ae07-71531815e0d7 · outbound

This paper cites Adding conditional control to text-to-image diffusion models.

Diffusing to Coordinate: Efficient Online Multi-Agent Diffusion Policies Adding conditional control to text-to-image diffusion models

Reference 71

Resolution
unresolved
no resolver link, observed 2026-08-02T22:04:27.830447Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T22:04:27.830447Z digest=sha256:0acf6774bf1feb8f874c4b0f82e2aa438e67b45c904d233817dd82fec2d6854d

Observation 9a75dd48-b868-470a-a3b0-232e0f0eb89b · outbound

This paper cites Scaling in-the-wild training for diffusion- based illumination harmonization and editing by imposing consistent light transport.

Diffusing to Coordinate: Efficient Online Multi-Agent Diffusion Policies Scaling in-the-wild training for diffusion- based illumination harmonization and editing by imposing consistent light transport

Reference 72

Resolution
unresolved
no resolver link, observed 2026-08-02T22:04:27.889476Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T22:04:27.889476Z digest=sha256:ba5bb9bcb0364b7d4d435cab5a1814bf776b33e3498be52e450ee7a277726411

Observation d35ad57d-25c7-4f7a-a625-0053afb43e14 · outbound

This paper cites Revisiting multi-agent world modeling from a diffusion-inspired perspective.

Diffusing to Coordinate: Efficient Online Multi-Agent Diffusion Policies Revisiting multi-agent world modeling from a diffusion-inspired perspective

Reference 73

Resolution
unresolved
no resolver link, observed 2026-08-02T22:04:27.965079Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T22:04:27.965079Z digest=sha256:171ff266a1ea388934f5a3ae02f31b2a6e6cbe1e87867fb68c8c80ff23aa2539

Observation 2297dc93-9b69-466c-b7a0-ba738651c7c7 · outbound

This paper cites Heterogeneous-agent reinforcement learning.Journal of Machine Learning Research, 25(32):1– 67, 2024.

Diffusing to Coordinate: Efficient Online Multi-Agent Diffusion Policies Heterogeneous-agent reinforcement learning.Journal of Machine Learning Research, 25(32):1– 67, 2024

Reference 74

Resolution
unresolved
no resolver link, observed 2026-08-02T22:04:28.039567Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T22:04:28.039567Z digest=sha256:90fb5cf4d4b4712bcf270343ec91dc41b0f538aaa16be16a356b85c926cc483e

Observation 9997169c-da64-4e85-9c82-6abccf050280 · outbound

This paper cites Smarts: An open-source scalable multi- agent rl training school for autonomous driving.

Diffusing to Coordinate: Efficient Online Multi-Agent Diffusion Policies Smarts: An open-source scalable multi- agent rl training school for autonomous driving

Reference 75

Resolution
unresolved
no resolver link, observed 2026-08-02T22:04:28.103641Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T22:04:28.103641Z digest=sha256:6386dc8fc3500611bd58ae6b811f72445126b4df1ad42a9fee226649bdca71b1

Observation 84cb4240-94d2-41e2-b980-4932a2befbdd · outbound

This paper cites Madiff: Offline multi-agent learning with diffusion models.Advances in Neural Information Processing Systems, 37:4177–4206, 2024.

Diffusing to Coordinate: Efficient Online Multi-Agent Diffusion Policies Madiff: Offline multi-agent learning with diffusion models.Advances in Neural Information Processing Systems, 37:4177–4206, 2024

Reference 76

Resolution
unresolved
no resolver link, observed 2026-08-02T22:04:28.180119Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T22:04:28.180119Z digest=sha256:75c3c369275f4b82eb2e4a776f6e9fa436ef59e1bd74f4122b17e57896d8eb0e

Observation 3b91f9b1-681f-45dd-9424-37451d130912 · outbound

This paper cites Maximum entropy inverse reinforcement learning.

Diffusing to Coordinate: Efficient Online Multi-Agent Diffusion Policies Maximum entropy inverse reinforcement learning

Reference 77

Resolution
unresolved
no resolver link, observed 2026-08-02T22:04:28.257554Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T22:04:28.257554Z digest=sha256:da8616e4b59b0944f26b2b45af7e4c61b1bffc56b71d9860c4352bdfd413636d

Observation 47fc1995-c07a-403a-a969-22a9f6efe9ee · outbound

This paper cites In spite of policy-centric ap- proaches, DIMA [73] uses diffusion models as environment dynamics to boost data efficiency.

Diffusing to Coordinate: Efficient Online Multi-Agent Diffusion Policies In spite of policy-centric ap- proaches, DIMA [73] uses diffusion models as environment dynamics to boost data efficiency

Reference 78

Resolution
unresolved
no resolver link, observed 2026-08-02T22:04:28.337206Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T22:04:28.337206Z digest=sha256:438921859bebd410bf3acc58de7a9d37316ef6ba66b8c3872de1d9b928d1ed84

Pith citing papers

Observation da7a6cb4-c6d0-4186-a5d2-0114d72cb8ec · inbound

Coordinated Diffusion: Generating Multi-Agent Behavior Without Multi-Agent Demonstrations cites this paper.

Coordinated Diffusion: Generating Multi-Agent Behavior Without Multi-Agent Demonstrations Diffusing to Coordinate: Efficient Online Multi-Agent Diffusion Policies

Reference 29

Resolution
verified exact
arxiv_id, observed 2026-06-11T02:08:43.143878Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-05-13T02:09:08.378144Z digest=sha256:81683db68ecf5dfbb0bc06c297c7c9b8c06096e2bbdd27981b483a1bc743de9e