Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-15T18:11:51.843712Z
Paper Citation Record · LEDGER
As of 21 August 2026, this Paper Citation Record lists 31 of 31 outbound references and 1 inbound Pith citation observation for arXiv:2507.18867.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-15T18:11:51.843712Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-21T06:32:19.484+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-06-30T22:11:35.277901Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z
31 of 31 outbound references displayed
External citation measurements
0
arxiv_reference, observed 2026-08-05T02:28:24.338817Z
Observation 86df1eb7-31ff-499f-8a05-7582d87f2d42 · outbound
Learning Individual Intrinsic Reward in Multi-Agent Reinforcement Learning via Incorporating Generalized Human Expertise An overview of recent progress in the study of distributed multi-agent coordination,
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 7aae10c2-ec9c-40f6-a210-59ccbd00f5a4 · outbound
Learning Individual Intrinsic Reward in Multi-Agent Reinforcement Learning via Incorporating Generalized Human Expertise Coordinated multi-agent reinforcement learning in networked distributed pomdps,
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 928bbe21-dea1-4800-92a2-d5c09fa0eed6 · outbound
Learning Individual Intrinsic Reward in Multi-Agent Reinforcement Learning via Incorporating Generalized Human Expertise Interpretation of neural networks is fragile,
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation f02dc4b7-aa54-4ba2-83bc-20b554f92d04 · outbound
Learning Individual Intrinsic Reward in Multi-Agent Reinforcement Learning via Incorporating Generalized Human Expertise Guided Deep Reinforcement Learning for Swarm Systems
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0ae5c23c-fd1b-49c2-8e86-f864dec45f2a · outbound
Learning Individual Intrinsic Reward in Multi-Agent Reinforcement Learning via Incorporating Generalized Human Expertise Q-value path decomposition for deep multiagent reinforce- ment learning,
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation f6ec4a4e-66db-4745-9627-d47e0086627b · outbound
Learning Individual Intrinsic Reward in Multi-Agent Reinforcement Learning via Incorporating Generalized Human Expertise RODE: Learning roles to decompose multi-agent tasks,
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation dc522ff2-f8d9-4f2d-8920-66be7dec5564 · outbound
Learning Individual Intrinsic Reward in Multi-Agent Reinforcement Learning via Incorporating Generalized Human Expertise ROMA: Multi-agent reinforcement learning with emergent roles,
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 17163ee2-9a0a-4b59-8fa9-6e68f2bfd633 · outbound
Learning Individual Intrinsic Reward in Multi-Agent Reinforcement Learning via Incorporating Generalized Human Expertise Value-decomposition networks for cooperative multi-agent learning based on team reward,
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 72450c64-c9f6-4d6b-a816-b14b5614cc75 · outbound
Learning Individual Intrinsic Reward in Multi-Agent Reinforcement Learning via Incorporating Generalized Human Expertise QMIX: Monotonic value function factorisation for deep multi-agent reinforcement learning,
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 7644ba24-bc25-4d19-b215-65910de58041 · outbound
Learning Individual Intrinsic Reward in Multi-Agent Reinforcement Learning via Incorporating Generalized Human Expertise QPLEX: Duplex dueling multi-agent Q-learning,
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation dd050ab3-74cd-4686-a4dd-6b607000b8a6 · outbound
Learning Individual Intrinsic Reward in Multi-Agent Reinforcement Learning via Incorporating Generalized Human Expertise Mixrts: Toward interpretable multi-agent reinforcement learning via mixing recurrent soft decision trees,
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7a55ea35-d8c7-4c25-b52b-79801420ac5c · outbound
Learning Individual Intrinsic Reward in Multi-Agent Reinforcement Learning via Incorporating Generalized Human Expertise NA 2Q: Neural attention additive model for interpretable multi-agent q-learning,
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation f4847e28-195d-45e3-b790-be2c8c1e4fec · outbound
Learning Individual Intrinsic Reward in Multi-Agent Reinforcement Learning via Incorporating Generalized Human Expertise A comprehensive survey of multiagent reinforcement learning,
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation cd4549f7-d7b7-467e-847c-81185da4cabc · outbound
Learning Individual Intrinsic Reward in Multi-Agent Reinforcement Learning via Incorporating Generalized Human Expertise Dynamic agent-based reward shaping for multi-agent systems,
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 135661cd-6a0e-4ecf-9350-bc2dede56e62 · outbound
Learning Individual Intrinsic Reward in Multi-Agent Reinforcement Learning via Incorporating Generalized Human Expertise Deep Multiagent Reinforcement Learning: Challenges and Directions
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c2afbab4-841b-45e2-a6c2-6f345f056307 · outbound
Learning Individual Intrinsic Reward in Multi-Agent Reinforcement Learning via Incorporating Generalized Human Expertise Cooperative exploration for multi-agent deep reinforcement learning,
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation df13e225-1866-4ce0-8d4a-c80628d35866 · outbound
Learning Individual Intrinsic Reward in Multi-Agent Reinforcement Learning via Incorporating Generalized Human Expertise Maven: Multi-agent variational exploration,
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation cb4634f4-e8f4-4aa7-9e1b-188afe0bbfb8 · outbound
Learning Individual Intrinsic Reward in Multi-Agent Reinforcement Learning via Incorporating Generalized Human Expertise Liir: Learning individual intrinsic reward in multi-agent reinforcement learning,
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation e152f0ad-e6b0-4619-be9d-de4d721080e5 · outbound
Learning Individual Intrinsic Reward in Multi-Agent Reinforcement Learning via Incorporating Generalized Human Expertise MASER: Multi-agent reinforcement learning with subgoals generated from experience replay buffer,
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 4c71277c-3f95-4217-ad70-7494054548b5 · outbound
Learning Individual Intrinsic Reward in Multi-Agent Reinforcement Learning via Incorporating Generalized Human Expertise Individual reward assisted multi-agent reinforcement learning,
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 0bd24b21-7f41-4da1-84a3-ca66dac89248 · outbound
Learning Individual Intrinsic Reward in Multi-Agent Reinforcement Learning via Incorporating Generalized Human Expertise Haven: hierarchical cooperative multi-agent reinforcement learning with dual coordination mechanism,
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 6b51b7fa-82ed-4b49-a18c-f78853a97a1d · outbound
Learning Individual Intrinsic Reward in Multi-Agent Reinforcement Learning via Incorporating Generalized Human Expertise Hierarchical Reinforcement Learning in StarCraft II with Human Expertise in Subgoals Selection
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation b0495970-2ab0-467c-9552-02fae64ee742 · outbound
Learning Individual Intrinsic Reward in Multi-Agent Reinforcement Learning via Incorporating Generalized Human Expertise Dl2: training and querying neural networks with logic,
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 47143ef4-d7de-4bbb-96df-ecc229f0b9ad · outbound
Learning Individual Intrinsic Reward in Multi-Agent Reinforcement Learning via Incorporating Generalized Human Expertise Rule-based reinforce- ment learning for efficient robot navigation with space reduction,
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation ff1811f7-7cea-4374-9ce0-e181e1928c31 · outbound
Learning Individual Intrinsic Reward in Multi-Agent Reinforcement Learning via Incorporating Generalized Human Expertise Extracting decision tree from trained deep reinforcement learning in traffic signal control,
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 3139a28d-2881-4a15-adec-4878867ed2ae · outbound
Learning Individual Intrinsic Reward in Multi-Agent Reinforcement Learning via Incorporating Generalized Human Expertise KoGuN: Accelerating Deep Reinforcement Learning via Integrating Human Suboptimal Knowledge
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 81a19c4a-408d-451b-b086-272ac28135e1 · outbound
Learning Individual Intrinsic Reward in Multi-Agent Reinforcement Learning via Incorporating Generalized Human Expertise QTRAN: Learning to factorize with transformation for cooperative multi-agent reinforcement learning,
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation c6b5f62d-b164-4fc5-949d-95b81035ffcf · outbound
Learning Individual Intrinsic Reward in Multi-Agent Reinforcement Learning via Incorporating Generalized Human Expertise Exploration with unreliable intrinsic reward in multi-agent reinforcement learning,
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 4e73d029-9e32-4d44-a71f-1a45254f122b · outbound
Learning Individual Intrinsic Reward in Multi-Agent Reinforcement Learning via Incorporating Generalized Human Expertise Discovering generalizable multi-agent coordination skills from multi-task offline data,
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation b16e8539-3a27-4d54-afe7-78d763cbaf48 · outbound
Learning Individual Intrinsic Reward in Multi-Agent Reinforcement Learning via Incorporating Generalized Human Expertise Shared experience actor- critic for multi-agent reinforcement learning,
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 2f20e365-ddce-4432-a5f0-8a1708048f2e · outbound
Learning Individual Intrinsic Reward in Multi-Agent Reinforcement Learning via Incorporating Generalized Human Expertise The StarCraft Multi-Agent Challenge,
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 4595a430-4bdd-4eb1-a04b-a8175adb6fd6 · inbound
Robust Instruction Compliance in Cooperative Multi-Agent Reinforcement Learning Learning Individual Intrinsic Reward in Multi-Agent Reinforcement Learning via Incorporating Generalized Human Expertise
Reference 64
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.