Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-05-23T04:16:54.807723Z
Paper Citation Record · LEDGER
As of 10 August 2026, this Paper Citation Record lists 29 of 29 outbound references and 0 inbound Pith citation observations for arXiv:2502.03506.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-05-23T04:16:54.807723Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
29 of 29 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 8d4aad54-5a9c-494e-81d3-b7f10e0e85e7 · outbound
Optimistic {\epsilon}-Greedy Exploration for Cooperative Multi-Agent Reinforcement Learning Exploration with Unreliable Intrinsic Reward in Multi-Agent Reinforcement Learning
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 28fa6271-c188-44fe-a5cc-1d07c8c45d4a · outbound
Optimistic {\epsilon}-Greedy Exploration for Cooperative Multi-Agent Reinforcement Learning The dynamics of reinforcement learning in cooperative multiagent systems
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 440c6bfa-76f0-4343-82ec-77bd178c0ebd · outbound
Optimistic {\epsilon}-Greedy Exploration for Cooperative Multi-Agent Reinforcement Learning Stabilising expe- rience replay for deep multi-agent reinforcement learning
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 3e3b0115-ad2a-446b-9b21-6580b0aea4c6 · outbound
Optimistic {\epsilon}-Greedy Exploration for Cooperative Multi-Agent Reinforcement Learning Counterfactual multi-agent policy gradients
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation ccf0af2c-98b7-43a9-9fd1-183ba17c465e · outbound
Optimistic {\epsilon}-Greedy Exploration for Cooperative Multi-Agent Reinforcement Learning Cirs: Bursting filter bubbles by counterfactual interactive recommender system
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 503cfa3a-9b97-4e47-be70-8eec28cc4376 · outbound
Optimistic {\epsilon}-Greedy Exploration for Cooperative Multi-Agent Reinforcement Learning Sampling efficient deep reinforcement learning through preference- guided stochastic exploration
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation eb549330-d5c8-4138-9a51-50d80196d2d5 · outbound
Optimistic {\epsilon}-Greedy Exploration for Cooperative Multi-Agent Reinforcement Learning Actor- attention-critic for multi-agent reinforcement learning
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 2fff0a42-23cb-4797-acc6-6e8a5edbce07 · outbound
Optimistic {\epsilon}-Greedy Exploration for Cooperative Multi-Agent Reinforcement Learning A Maximum Mutual Information Framework for Multi-Agent Reinforcement Learning
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 892fa9bd-5464-416a-969e-728259483c98 · outbound
Optimistic {\epsilon}-Greedy Exploration for Cooperative Multi-Agent Reinforcement Learning Multi-agent reinforcement learning for traffic signal control: A cooperative approach
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation aae924d3-e707-4255-91fb-fc0d536bdac8 · outbound
Optimistic {\epsilon}-Greedy Exploration for Cooperative Multi-Agent Reinforcement Learning A unified game- theoretic approach to multiagent reinforcement learning
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 5f3e2173-aa79-455d-8cf5-0d511a204eec · outbound
Optimistic {\epsilon}-Greedy Exploration for Cooperative Multi-Agent Reinforcement Learning Optimistic value instructors for co- operative multi-agent reinforcement learning
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation baf7b414-d434-4504-abaf-5e089e76c9e6 · outbound
Optimistic {\epsilon}-Greedy Exploration for Cooperative Multi-Agent Reinforcement Learning Markov games as a framework for multi-agent reinforcement learning
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 5debd818-3b4c-473a-aac7-db4c96854143 · outbound
Optimistic {\epsilon}-Greedy Exploration for Cooperative Multi-Agent Reinforcement Learning Cooperative exploration for multi-agent deep reinforcement learning
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation aed44080-2486-40ba-ab85-1b01755c96a4 · outbound
Optimistic {\epsilon}-Greedy Exploration for Cooperative Multi-Agent Reinforcement Learning Multi- agent actor-critic for mixed cooperative-competitive envi- ronments
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 598393ff-2c5b-4228-b744-c825964102a1 · outbound
Optimistic {\epsilon}-Greedy Exploration for Cooperative Multi-Agent Reinforcement Learning Likelihood Quantile Networks for Coordinating Multi-Agent Reinforcement Learning
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation d1922028-8854-4da9-97d6-ebce5426a694 · outbound
Optimistic {\epsilon}-Greedy Exploration for Cooperative Multi-Agent Reinforcement Learning Benchmarking Multi-Agent Deep Reinforcement Learning Algorithms in Cooperative Tasks
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 7382ba68-2bbd-47a7-8697-4572e85afe5e · outbound
Optimistic {\epsilon}-Greedy Exploration for Cooperative Multi-Agent Reinforcement Learning Residual q-networks for value function factorizing in multiagent reinforcement learning
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 5e9f1047-0c1b-4787-80cc-b805414b23a1 · outbound
Optimistic {\epsilon}-Greedy Exploration for Cooperative Multi-Agent Reinforcement Learning The StarCraft Multi-Agent Challenge
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 3bb3f69d-8882-48fa-9ba7-10376295c472 · outbound
Optimistic {\epsilon}-Greedy Exploration for Cooperative Multi-Agent Reinforcement Learning Safe, Multi-Agent, Reinforcement Learning for Autonomous Driving
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 6cf9922e-8d22-403d-a9fa-cc4706fea37d · outbound
Optimistic {\epsilon}-Greedy Exploration for Cooperative Multi-Agent Reinforcement Learning Resq: A residual q function-based approach for multi-agent reinforcement learning value factoriza- tion
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 78741196-64ef-4eec-86db-bf3de7557ce7 · outbound
Optimistic {\epsilon}-Greedy Exploration for Cooperative Multi-Agent Reinforcement Learning Qtran: Learn- ing to factorize with transformation for cooperative multi- agent reinforcement learning
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 58ef8e98-0c50-44e1-ab79-74985af328e2 · outbound
Optimistic {\epsilon}-Greedy Exploration for Cooperative Multi-Agent Reinforcement Learning Dfac framework: Factorizing the value function via quantile mixture for multi-agent distribu- tional q-learning
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 8ffdbc16-ac46-4810-a962-23a104ba1e77 · outbound
Optimistic {\epsilon}-Greedy Exploration for Cooperative Multi-Agent Reinforcement Learning Value-Decomposition Networks For Cooperative Multi-Agent Learning
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 33a79463-f57b-475e-b3ad-20b001c51f8c · outbound
Optimistic {\epsilon}-Greedy Exploration for Cooperative Multi-Agent Reinforcement Learning Influence-Based Multi-Agent Exploration
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 152f135b-43f8-4049-9623-59100e86958d · outbound
Optimistic {\epsilon}-Greedy Exploration for Cooperative Multi-Agent Reinforcement Learning QPLEX: Duplex Dueling Multi-Agent Q-Learning
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 47e915b9-1ee9-4749-a7a4-76832870e21d · outbound
Optimistic {\epsilon}-Greedy Exploration for Cooperative Multi-Agent Reinforcement Learning En- hancing collaboration in multi-agent reinforcement learn- ing with correlated trajectories
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 3bf41391-8c21-4374-937f-80f6536a38ff · outbound
Optimistic {\epsilon}-Greedy Exploration for Cooperative Multi-Agent Reinforcement Learning Fully decentral- ized multi-agent reinforcement learning with networked agents
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 871cbf58-fcba-4061-823a-0f2f95f7812e · outbound
Optimistic {\epsilon}-Greedy Exploration for Cooperative Multi-Agent Reinforcement Learning Condi- tionally optimistic exploration for cooperative deep multi- agent reinforcement learning
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation a4484c9f-959e-48e1-9f12-df838671ca77 · outbound
Optimistic {\epsilon}-Greedy Exploration for Cooperative Multi-Agent Reinforcement Learning Qdap: Downsizing adaptive policy for cooperative multi-agent reinforcement learning
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
No inbound Pith citation observations are available.