Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T13:57:40.245927Z
Paper Citation Record · LEDGER
As of 14 August 2026, this Paper Citation Record lists 32 of 32 outbound references and 0 inbound Pith citation observations for arXiv:2505.20579.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T13:57:40.245927Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-13T06:32:02.005865+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
32 of 32 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 0785d4bb-e775-4cfb-8e2c-d5f46fb58dfd · outbound
The challenge of hidden gifts in multi-agent reinforcement learning LOQA : Learning with opponent q-learning awareness
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 65b50dc3-88b8-429a-8651-2d1ae047486d · outbound
The challenge of hidden gifts in multi-agent reinforcement learning Unifying temporal and structural credit assignment problems
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 45148112-d0c6-4aa3-becf-9b96e2203caf · outbound
The challenge of hidden gifts in multi-agent reinforcement learning Understanding the impact of entropy on policy optimization
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 717fd1f9-4ef1-4ca7-92c3-7639a377b6b5 · outbound
The challenge of hidden gifts in multi-agent reinforcement learning Effective choice in the prisoner's dilemma
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 545c2036-bae3-4537-a921-a7b2ee20f4d5 · outbound
The challenge of hidden gifts in multi-agent reinforcement learning Manitokanac
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 878a3ef2-3b7d-4462-b2c7-3f02961b23a5 · outbound
The challenge of hidden gifts in multi-agent reinforcement learning The theory of dynamic programming
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fa69cbd5-ecc0-40ef-98f3-07dacafc609b · outbound
The challenge of hidden gifts in multi-agent reinforcement learning Prisoner's dilemma; a study in conflict and cooperation
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation fecfc178-cc40-4f6c-ad19-b731c3caee36 · outbound
The challenge of hidden gifts in multi-agent reinforcement learning Decision transformer: Reinforcement learning via sequence modeling
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c395c7bc-213f-4ad1-ab39-50dc2bea790a · outbound
The challenge of hidden gifts in multi-agent reinforcement learning Minigrid & miniworld: Modular & customizable reinforcement learning environments for goal-oriented tasks
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 440b0c45-833b-4302-9b72-75deec581e6f · outbound
The challenge of hidden gifts in multi-agent reinforcement learning Learning phrase representations using RNN encoder -- decoder for statistical machine translation
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1429d519-a76c-4b38-8964-b27f8e51bc77 · outbound
The challenge of hidden gifts in multi-agent reinforcement learning Hypothetical minds: Scaffolding theory of mind for multi-agent tasks with large language models
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation ee237fee-8184-496b-b614-bec4548cb90b · outbound
The challenge of hidden gifts in multi-agent reinforcement learning Maximum entropy RL (provably) solves some robust RL problems
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation edcda3d2-97cb-44e2-b238-7b58dad96676 · outbound
The challenge of hidden gifts in multi-agent reinforcement learning Counterfactual multi-agent policy gradients
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1d218df6-f17b-4151-b312-7a51599cfd3e · outbound
The challenge of hidden gifts in multi-agent reinforcement learning Learning with Opponent-Learning Awareness
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fa5fa247-7d8a-4ce1-90c4-4e7525eabf33 · outbound
The challenge of hidden gifts in multi-agent reinforcement learning Structural credit assignment in neural networks using reinforcement learning
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 23447fde-0e28-4662-8aaf-02fb51e13281 · outbound
The challenge of hidden gifts in multi-agent reinforcement learning Soft actor-critic: Off-policy maximum entropy deep reinforcement learning with a stochastic actor
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c9e0a6a7-9510-41ae-9122-827df185a5cf · outbound
The challenge of hidden gifts in multi-agent reinforcement learning Optimizing agent behavior over long time scales by transporting value
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 0e6b197b-fa05-4ba0-96bc-c5aeff909f22 · outbound
The challenge of hidden gifts in multi-agent reinforcement learning Stateful active facilitator: Coordination and environmental heterogeneity in cooperative multi-agent reinforcement learning
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 38f3c47b-f42f-45cf-add0-5cde06837423 · outbound
The challenge of hidden gifts in multi-agent reinforcement learning Maven: Multi-agent variational exploration
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5778cf27-9108-4fa6-9e27-766e4d1956f5 · outbound
The challenge of hidden gifts in multi-agent reinforcement learning Multi-agent cooperation through learning-aware policy gradients
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 86740e3a-9d60-4325-aeb7-1832e2994eb9 · outbound
The challenge of hidden gifts in multi-agent reinforcement learning Equilibrium points in n-person games
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5cf85eeb-2d4b-4a4e-b747-7bc456e0db6b · outbound
The challenge of hidden gifts in multi-agent reinforcement learning When do transformers shine in rl? decoupling memory from credit assignment
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 55335269-7cf6-46a5-9871-7febce7d37f4 · outbound
The challenge of hidden gifts in multi-agent reinforcement learning Monotonic value function factorisation for deep multi-agent reinforcement learning
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 451b9fd3-8fe7-4219-8eaa-1f1103f85942 · outbound
The challenge of hidden gifts in multi-agent reinforcement learning Proximal Policy Optimization Algorithms
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e3603498-d8ec-4585-ae22-691096eaa45e · outbound
The challenge of hidden gifts in multi-agent reinforcement learning Agent-Time Attention for Sparse Rewards Multi-Agent Reinforcement Learning
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation a839fd47-f4ca-41e4-84a2-aa7124976a93 · outbound
The challenge of hidden gifts in multi-agent reinforcement learning Qtran: Learning to factorize with transformation for cooperative multi-agent reinforcement learning
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation db7404eb-e5c0-4b34-a830-334337701ea0 · outbound
The challenge of hidden gifts in multi-agent reinforcement learning PufferLib: Making Reinforcement Learning Libraries and Environments Play Nice
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 91fd2b73-84ff-4050-9654-cc1a7d4f5e5a · outbound
The challenge of hidden gifts in multi-agent reinforcement learning Value-decomposition networks for cooperative multi-agent learning
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 6588f00d-0d2f-4415-a0a9-d4590e78ca2a · outbound
The challenge of hidden gifts in multi-agent reinforcement learning Policy gradient methods for reinforcement learning with function approximation
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 32731990-74ff-40fb-8c52-8d65ba72f489 · outbound
The challenge of hidden gifts in multi-agent reinforcement learning Learning sequences of actions in collectives of autonomous agents
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 4873364c-44db-423f-8ed7-6303da79cba4 · outbound
The challenge of hidden gifts in multi-agent reinforcement learning Cola: consistent learning with opponent-learning awareness
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 4e88e8fe-79cc-47f6-9a09-8e48e4f66aa1 · outbound
The challenge of hidden gifts in multi-agent reinforcement learning The surprising effectiveness of ppo in cooperative multi-agent games
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
No inbound Pith citation observations are available.