Pith. sign in

Paper Citation Record · LEDGER

Aggregate in the Advantage, Not the Ratio: A Canonical-Form Analysis of Cooperative Multi-Agent Policy Optimization

As of 10 August 2026, this Paper Citation Record lists 12 of 12 outbound references and 0 inbound Pith citation observations for arXiv:2607.17924.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2607.17924 v1

Coverage vector

measured 12 of 12 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-01T16:42:04.090374Z

measured 12 of 12 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

12 of 12 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved12
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation dfb80bfe-04ff-4c44-bc48-6d234cb14935 · outbound

This paper cites Is Independent Learning All You Need in the StarCraft Multi-Agent Challenge?.

Aggregate in the Advantage, Not the Ratio: A Canonical-Form Analysis of Cooperative Multi-Agent Policy Optimization Is Independent Learning All You Need in the StarCraft Multi-Agent Challenge?

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-01T16:42:02.462350Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T16:42:02.462350Z digest=sha256:c863bed16240e51b0b8dfb5bf069c38fe857749706ad937b141cdfc67dc9378a

Observation e86fa2a0-4660-40d3-8c4f-a290712c33d6 · outbound

This paper cites Bridging MARL to SARL: An Order-Independent Multi-Agent Transformer via Latent Consensus.

Aggregate in the Advantage, Not the Ratio: A Canonical-Form Analysis of Cooperative Multi-Agent Policy Optimization Bridging MARL to SARL: An Order-Independent Multi-Agent Transformer via Latent Consensus

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-01T16:42:03.341501Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T16:42:03.341501Z digest=sha256:83b241b93b708cfff5c7f158d9511df4a999ea6f893b78582ea9737e004ef18b

Observation 7f9ffecb-a468-438e-b0bc-786abaf1f2c7 · outbound

This paper cites an unresolved cited work.

Aggregate in the Advantage, Not the Ratio: A Canonical-Form Analysis of Cooperative Multi-Agent Policy Optimization Unresolved cited work

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-01T16:42:03.917531Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T16:42:03.917531Z digest=sha256:acea050891db800f0877afd72cd39cecfa5b2dac0c431be3a62e842ccbc36d0a

Observation 3d652191-670e-4ac6-beae-fb61a0474d0f · outbound

This paper cites an unresolved cited work.

Aggregate in the Advantage, Not the Ratio: A Canonical-Form Analysis of Cooperative Multi-Agent Policy Optimization Unresolved cited work

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-01T16:42:04.090374Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T16:42:04.090374Z digest=sha256:d847f7c7f4164c77ee8cb46ffa2e9f5640f2b2089231144feef38ab78a58ebdb

Observation 70fece1c-6a00-458e-a4cf-a8cf22f55e76 · outbound

This paper cites High-Dimensional Continuous Control Using Generalized Advantage Estimation.

Aggregate in the Advantage, Not the Ratio: A Canonical-Form Analysis of Cooperative Multi-Agent Policy Optimization High-Dimensional Continuous Control Using Generalized Advantage Estimation

Reference 1994

Resolution
unresolved
no resolver link, observed 2026-08-01T16:42:03.179272Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T16:42:03.179272Z digest=sha256:132ffc99ffaf4f81d4be9c873ef6adfb1034986655df45f99b8372e74ca108e1

Observation 98b8ac4a-bf9b-41ac-bb3d-c29bff3ef3e2 · outbound

This paper cites an unresolved cited work.

Aggregate in the Advantage, Not the Ratio: A Canonical-Form Analysis of Cooperative Multi-Agent Policy Optimization Unresolved cited work

Reference 2011

Resolution
unresolved
no resolver link, observed 2026-08-01T16:42:03.687268Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T16:42:03.687268Z digest=sha256:dc9e96a8f0e2ee74ed8285e53d484e8c5dfa6830088b32c59e45367891e82f54

Observation 28c4cdc4-dd85-4fcf-88e9-e25e415ddba4 · outbound

This paper cites A Comprehensive Survey on Multi-Agent Cooperative Decision-Making: Scenarios, Approaches, Challenges and Perspectives.

Aggregate in the Advantage, Not the Ratio: A Canonical-Form Analysis of Cooperative Multi-Agent Policy Optimization A Comprehensive Survey on Multi-Agent Cooperative Decision-Making: Scenarios, Approaches, Challenges and Perspectives

Reference 2018

Resolution
unresolved
no resolver link, observed 2026-08-01T16:42:02.557564Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T16:42:02.557564Z digest=sha256:375ba363a44059caaf76e8db8aea2e9107a36f4d10d80bec68e88f6565353837

Observation a830c2a8-6882-4e8d-9f3e-8dd3eb6a4478 · outbound

This paper cites Sumo–simulation of urban mobility: an overview.

Aggregate in the Advantage, Not the Ratio: A Canonical-Form Analysis of Cooperative Multi-Agent Policy Optimization Sumo–simulation of urban mobility: an overview

Reference 2019

Resolution
unresolved
no resolver link, observed 2026-08-01T16:42:02.386172Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T16:42:02.386172Z digest=sha256:9f12e5c8c27a12c0810053651d96bbe8acc56f81f94fdf3b851c2ea3f51b2225

Observation 9cdcdc7e-db66-4795-8f60-5af32579eef5 · outbound

This paper cites Markov games as a framework for multi-agent reinforcement learning.

Aggregate in the Advantage, Not the Ratio: A Canonical-Form Analysis of Cooperative Multi-Agent Policy Optimization Markov games as a framework for multi-agent reinforcement learning

Reference 2022

Resolution
unresolved
no resolver link, observed 2026-08-01T16:42:02.957872Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T16:42:02.957872Z digest=sha256:5d7b1405f0385a2166c0a49e9a18a9d563e68ccfbdd87d7e69fec54a84215c43

Observation 647be801-6234-4bea-b8bc-1ab307dcc323 · outbound

This paper cites 13 B.2 Proof of Lemma 2 (multiplicative variance).

Aggregate in the Advantage, Not the Ratio: A Canonical-Form Analysis of Cooperative Multi-Agent Policy Optimization 13 B.2 Proof of Lemma 2 (multiplicative variance)

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-08-01T16:42:03.520863Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T16:42:03.520863Z digest=sha256:acc14725eeefa960b0e8ae1e1b91d1dee41e351a8a79158165b93c1a6ae4fd05

Observation 1326ed37-126f-4953-9039-bc2868288590 · outbound

This paper cites Generalized per-agent advantage estimation for multi-agent policy optimization.arXiv preprint arXiv:2603.02654,.

Aggregate in the Advantage, Not the Ratio: A Canonical-Form Analysis of Cooperative Multi-Agent Policy Optimization Generalized per-agent advantage estimation for multi-agent policy optimization.arXiv preprint arXiv:2603.02654,

Reference 2025

Resolution
unresolved
no resolver link, observed 2026-08-01T16:42:02.678810Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T16:42:02.678810Z digest=sha256:55083df861100f557b974e7ddaf94d5776ce6b12d88d3960c4a4423baf5b861e

Observation 45d08777-a3d6-49ab-accc-75a350072ee6 · outbound

This paper cites Trust region policy optimisation in multi-agent reinforcement learning.

Aggregate in the Advantage, Not the Ratio: A Canonical-Form Analysis of Cooperative Multi-Agent Policy Optimization Trust region policy optimisation in multi-agent reinforcement learning

Reference 2026

Resolution
unresolved
no resolver link, observed 2026-08-01T16:42:02.792432Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T16:42:02.792432Z digest=sha256:ad4714307b446a86d6c6bae3d10060bad7a6972c18465e99f7fe464c6181a1ee

Pith citing papers

No inbound Pith citation observations are available.