Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-09T01:03:57.025841Z
Paper Citation Record · LEDGER
As of 9 August 2026, this Paper Citation Record lists 15 of 15 outbound references and 2 inbound Pith citation observations for arXiv:2502.03723.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-09T01:03:57.025841Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-04T07:56:46.797195Z
A source-named dated measurement, never combined with another source.
Source: cited_works
15 of 15 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 85dd0847-b73e-4d6f-b06c-61836ec1bfe7 · outbound
Speaking the Language of Teamwork: LLM-Guided Credit Assignment in Multi-Agent Reinforcement Learning GPT-4 Technical Report
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation da73d674-217f-498e-b824-a679cfa9a230 · outbound
Speaking the Language of Teamwork: LLM-Guided Credit Assignment in Multi-Agent Reinforcement Learning RLAIF vs. RLHF: Scaling Reinforcement Learning from Human Feedback with AI Feedback
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 67a4b0f7-931d-4a38-82ba-3573b732027a · outbound
Speaking the Language of Teamwork: LLM-Guided Credit Assignment in Multi-Agent Reinforcement Learning PEBBLE: Feedback-Efficient Interactive Reinforcement Learning via Relabeling Experience and Unsupervised Pre-training
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f4c0a592-b42c-4738-a2c3-0f7e97a1021d · outbound
Speaking the Language of Teamwork: LLM-Guided Credit Assignment in Multi-Agent Reinforcement Learning Value-Decomposition Multi-Agent Actor-Critics
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 07a7d512-4dd4-49e0-a511-8fea0fc40018 · outbound
Speaking the Language of Teamwork: LLM-Guided Credit Assignment in Multi-Agent Reinforcement Learning A Minimaximalist Approach to Reinforcement Learning from Human Feedback
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a4bd5be9-8ec5-452e-9614-099159e616da · outbound
Speaking the Language of Teamwork: LLM-Guided Credit Assignment in Multi-Agent Reinforcement Learning QPLEX: Duplex Dueling Multi-Agent Q-Learning
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f1e0ea21-0d5d-40ff-9d48-64a1ab8391f9 · outbound
Speaking the Language of Teamwork: LLM-Guided Credit Assignment in Multi-Agent Reinforcement Learning The Surprising Effectiveness of PPO in Cooperative, Multi-Agent Games
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1a12b5d7-adb3-437f-803b-e196361904d1 · outbound
Speaking the Language of Teamwork: LLM-Guided Credit Assignment in Multi-Agent Reinforcement Learning Fine-Tuning Language Models from Human Preferences
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c015bae3-3cdc-4418-9e2b-b7d71171aa35 · outbound
Speaking the Language of Teamwork: LLM-Guided Credit Assignment in Multi-Agent Reinforcement Learning Safe, Multi-Agent, Reinforcement Learning for Autonomous Driving
Reference 2017
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 71e1e9c4-e878-4489-a434-6e91cfa7c096 · outbound
Speaking the Language of Teamwork: LLM-Guided Credit Assignment in Multi-Agent Reinforcement Learning M., Stepputtis, S., Campbell, J., and Sycara, K
Reference 2018
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation fd87cffd-a3a8-4b0c-9528-235a4f84bdcb · outbound
Speaking the Language of Teamwork: LLM-Guided Credit Assignment in Multi-Agent Reinforcement Learning Unresolved cited work
Reference 2020
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 081a3b8a-f58e-4007-8964-81aa60c348bd · outbound
Speaking the Language of Teamwork: LLM-Guided Credit Assignment in Multi-Agent Reinforcement Learning LLaMA: Open and Efficient Foundation Language Models
Reference 2021
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b66be455-9538-49cb-a6d0-4c0848d20552 · outbound
Speaking the Language of Teamwork: LLM-Guided Credit Assignment in Multi-Agent Reinforcement Learning Multi-Agent Reinforcement Learning is a Sequence Modeling Problem
Reference 2022
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 8ea196f9-74d8-4fd3-b484-b654409ff70e · outbound
Speaking the Language of Teamwork: LLM-Guided Credit Assignment in Multi-Agent Reinforcement Learning A variational approach to mutual information-based coordination for multi-agent reinforcement learning
Reference 2023
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation cff30953-901f-495c-928c-8f3660cacebc · outbound
Speaking the Language of Teamwork: LLM-Guided Credit Assignment in Multi-Agent Reinforcement Learning Assigning Credit with Partial Reward Decoupling in Multi-Agent Proximal Policy Optimization
Reference 2024
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 61540494-045a-46b4-b6e9-baa27c71ad69 · inbound
MASPRM: Multi-Agent System Process Reward Model Speaking the Language of Teamwork: LLM-Guided Credit Assignment in Multi-Agent Reinforcement Learning
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation be47df70-0bf2-4a7c-ad41-141f96d55027 · inbound
MARS-RA: Rank Aggregation for Credit Assignment via Multimodal Comparisons in Embodied Multi-Agent Cooperation Speaking the Language of Teamwork: LLM-Guided Credit Assignment in Multi-Agent Reinforcement Learning
Reference 81
Source-reported events for the cited work
Unavailable: canonical work link unavailable.