Pith. sign in

Paper Citation Record · LEDGER

Scalable Policy Optimization for Networked Multi-Agent Reinforcement Learning with Continuous State-Action Spaces

As of 9 August 2026, this Paper Citation Record lists 34 of 34 outbound references and 0 inbound Pith citation observations for arXiv:2607.18554.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2607.18554 v1

Coverage vector

measured 34 of 34 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-01T15:07:45.671697Z

measured 34 of 34 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

34 of 34 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved34
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation db074900-9149-4f9e-9ff3-efba1aaf7f94 · outbound

This paper cites Multi-agent reinforcement learning: A selective overview of theories and algorithms,.

Scalable Policy Optimization for Networked Multi-Agent Reinforcement Learning with Continuous State-Action Spaces Multi-agent reinforcement learning: A selective overview of theories and algorithms,

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-01T15:07:44.131604Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T15:07:44.131604Z digest=sha256:cd13b62942c84a1ea0d82e6fe8cd487bf900ab49c8b3c7831101b43173b9c5e0

Observation 4d0640fd-9165-401e-a914-f4e93c376b17 · outbound

This paper cites Stability constrained reinforcement learning for decentralized real-time voltage control,.

Scalable Policy Optimization for Networked Multi-Agent Reinforcement Learning with Continuous State-Action Spaces Stability constrained reinforcement learning for decentralized real-time voltage control,

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-01T15:07:44.242509Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T15:07:44.242509Z digest=sha256:0426032f9538cf5a0d9e46104bf55e60934b4a1cee1bc8c7dbeb3749af511b95

Observation 98d8d835-e27f-40d5-9f1e-f373ce4ec53b · outbound

This paper cites Scalable reinforcement learning of localized policies for multi-agent networked systems,.

Scalable Policy Optimization for Networked Multi-Agent Reinforcement Learning with Continuous State-Action Spaces Scalable reinforcement learning of localized policies for multi-agent networked systems,

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-01T15:07:44.336581Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T15:07:44.336581Z digest=sha256:2a49a08eeecae5bb8039bdea539dd700b2ad068d0fb6bfb5bc966f9aba6aebf7

Observation 39abe093-75f4-43d7-bcd1-2ed504277f55 · outbound

This paper cites Scalable reinforcement learning for multiagent networked sys- tems,.

Scalable Policy Optimization for Networked Multi-Agent Reinforcement Learning with Continuous State-Action Spaces Scalable reinforcement learning for multiagent networked sys- tems,

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-01T15:07:44.428470Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T15:07:44.428470Z digest=sha256:5ddc3d7b906573b35b2ae86592753d05f2f6191c183f3052fd076d83e6823090

Observation 9b0bebf4-cc51-4cac-8dcc-91045062ae00 · outbound

This paper cites Multi-agent reinforcement learning in stochastic networked systems,.

Scalable Policy Optimization for Networked Multi-Agent Reinforcement Learning with Continuous State-Action Spaces Multi-agent reinforcement learning in stochastic networked systems,

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-01T15:07:44.504402Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T15:07:44.504402Z digest=sha256:22be2dec0c077a7db7c13327be5cfe6c79ffe7a49e20affb2bb47f212ae40633

Observation 26d11f3e-74f4-45f7-b721-4fbdb9af915f · outbound

This paper cites Global convergence of localized policy iteration in networked multi-agent reinforcement learning,.

Scalable Policy Optimization for Networked Multi-Agent Reinforcement Learning with Continuous State-Action Spaces Global convergence of localized policy iteration in networked multi-agent reinforcement learning,

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-01T15:07:44.587308Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T15:07:44.587308Z digest=sha256:116bf54c09032ef9eaa131de356b3d4e2fc33d40d647d16c6205a5489d4db411

Observation a5fd7b01-b3ea-470a-983e-9fb19504aaac · outbound

This paper cites Fully decentralized multi-agent reinforcement learning with networked agents,.

Scalable Policy Optimization for Networked Multi-Agent Reinforcement Learning with Continuous State-Action Spaces Fully decentralized multi-agent reinforcement learning with networked agents,

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-01T15:07:44.672839Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T15:07:44.672839Z digest=sha256:300acbbe2589475a55a05c67918785a8c8947aeb4e4307077765baf41c2b0965

Observation 380f43a1-8d15-4e7e-a29a-f779e003e649 · outbound

This paper cites Finite-time analysis of dis- tributed td (0) with linear function approximation on multi-agent rein- forcement learning,.

Scalable Policy Optimization for Networked Multi-Agent Reinforcement Learning with Continuous State-Action Spaces Finite-time analysis of dis- tributed td (0) with linear function approximation on multi-agent rein- forcement learning,

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-01T15:07:44.749424Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T15:07:44.749424Z digest=sha256:080ca66d1aa8540089ec4f9212da0169140440ed22a61643468f8dada4dbe082

Observation 80fce664-357f-49ca-ac5e-0e9859a7e637 · outbound

This paper cites Decentralized online convex optimization in networked systems,.

Scalable Policy Optimization for Networked Multi-Agent Reinforcement Learning with Continuous State-Action Spaces Decentralized online convex optimization in networked systems,

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-01T15:07:44.816692Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T15:07:44.816692Z digest=sha256:b87b99494218449f6c0484956a3063a39172dcaf48241a549dc4435f4d3217e3

Observation ff360be4-3fe0-43b1-a6e6-7082e426bde6 · outbound

This paper cites Multi-agent Reinforcement Learning for Networked System Control.

Scalable Policy Optimization for Networked Multi-Agent Reinforcement Learning with Continuous State-Action Spaces Multi-agent Reinforcement Learning for Networked System Control

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-01T15:07:44.900715Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T15:07:44.900715Z digest=sha256:19876aded5199165cf8970b82e336618fa9ca37d21157e35532795fdce7f1499

Observation ff10ef29-0db5-445e-b3dd-d0e651b15034 · outbound

This paper cites Random features for large-scale kernel machines,.

Scalable Policy Optimization for Networked Multi-Agent Reinforcement Learning with Continuous State-Action Spaces Random features for large-scale kernel machines,

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-01T15:07:44.968846Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T15:07:44.968846Z digest=sha256:9d2c8743ab75df91a00fe5942ff0f83c621347034a463b1d41a5ba9ba54dd51d

Observation 45b6321f-8983-43f4-8e68-0aa60bec6e0b · outbound

This paper cites Random features for ker- nel approximation: A survey on algorithms, theory, and beyond,.

Scalable Policy Optimization for Networked Multi-Agent Reinforcement Learning with Continuous State-Action Spaces Random features for ker- nel approximation: A survey on algorithms, theory, and beyond,

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-01T15:07:45.043151Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T15:07:45.043151Z digest=sha256:060ec5d384b1ad076c385bd057f6fdf73315e8ad26cc7dd5d6d68445c78e24bd

Observation 0758319c-58e1-4bb2-ae17-8568332d969e · outbound

This paper cites Scalable spectral representations for multi-agent reinforcement learning in network MDPs.

Scalable Policy Optimization for Networked Multi-Agent Reinforcement Learning with Continuous State-Action Spaces Scalable spectral representations for multi-agent reinforcement learning in network MDPs

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-01T15:07:45.122994Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T15:07:45.122994Z digest=sha256:882b31e2a289b71c5cd1f3cfac38b3b9e6c7309cc439a73b4103723cdb28be60

Observation af3c1833-2880-4063-92fd-cfa01126defc · outbound

This paper cites Linear least-squares algorithms for temporal difference learning,.

Scalable Policy Optimization for Networked Multi-Agent Reinforcement Learning with Continuous State-Action Spaces Linear least-squares algorithms for temporal difference learning,

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-01T15:07:45.209477Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T15:07:45.209477Z digest=sha256:b80e073ed4aeabe43d150672e2e3e11642b77f43aca751b4c0f183350c77cfa1

Observation 820e9772-ffdf-4788-abfe-c477b2a7d58c · outbound

This paper cites Technical update: Least-squares temporal difference learning,.

Scalable Policy Optimization for Networked Multi-Agent Reinforcement Learning with Continuous State-Action Spaces Technical update: Least-squares temporal difference learning,

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-01T15:07:45.298257Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T15:07:45.298257Z digest=sha256:8a55da6c81abb0c927142b43b7d15b02ddce240cf16208caf1bfd10f6764691d

Observation 9c84af58-9a4d-4e35-8308-b8ae758afa4d · outbound

This paper cites Finite-sample analysis of least-squares policy iteration,.

Scalable Policy Optimization for Networked Multi-Agent Reinforcement Learning with Continuous State-Action Spaces Finite-sample analysis of least-squares policy iteration,

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-01T15:07:45.391730Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T15:07:45.391730Z digest=sha256:7acfbda634a6ac168b12f4a15a2a1d63c390cd128c6b282cd8998fe15c57827a

Observation bf0b56bf-d20b-4ebe-8711-5e4887d30db3 · outbound

This paper cites A finite time analysis of temporal difference learning with linear function approximation,.

Scalable Policy Optimization for Networked Multi-Agent Reinforcement Learning with Continuous State-Action Spaces A finite time analysis of temporal difference learning with linear function approximation,

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-01T15:07:45.455637Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T15:07:45.455637Z digest=sha256:96a4847e9e8d92a0bd77897ffb798372a8913c523bf53c1bf6fab176e2d74cc6

Observation fc623c16-c7c9-42c7-b88d-12c72fc7b6bf · outbound

This paper cites An introduction to matrix concentration inequalities,.

Scalable Policy Optimization for Networked Multi-Agent Reinforcement Learning with Continuous State-Action Spaces An introduction to matrix concentration inequalities,

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-01T15:07:45.493186Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T15:07:45.493186Z digest=sha256:50127ad3f3cabed4033ba1f34c4907bde1391417c5566753f75c3005e95f68a5

Observation 8a738021-d68d-4b74-aa2f-2c70b18a138e · outbound

This paper cites Vershynin,High-dimensional probability: An introduction with ap- plications in data science.

Scalable Policy Optimization for Networked Multi-Agent Reinforcement Learning with Continuous State-Action Spaces Vershynin,High-dimensional probability: An introduction with ap- plications in data science

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-01T15:07:45.534618Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T15:07:45.534618Z digest=sha256:6115b481b0a8be30b6048103ec8f497d202c99cd2a7d933a58a618b140ba7154

Observation 9912ec25-11ff-43f0-b13c-6c977f9c24e1 · outbound

This paper cites Optimum bounds for the distributions of martingales in banach spaces,.

Scalable Policy Optimization for Networked Multi-Agent Reinforcement Learning with Continuous State-Action Spaces Optimum bounds for the distributions of martingales in banach spaces,

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-01T15:07:45.609186Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T15:07:45.609186Z digest=sha256:2ee65b9037dc19ca8c8295cbd0d6b4ab925ba82d77c3dc8b6f504bb872f41ff4

Observation 02ba7837-d616-490e-971e-35b18668207b · outbound

This paper cites an unresolved cited work.

Scalable Policy Optimization for Networked Multi-Agent Reinforcement Learning with Continuous State-Action Spaces Unresolved cited work

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-01T15:07:45.625579Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T15:07:45.625579Z digest=sha256:4dbd63ae46d64cf33c374b7ff76865004d968b32f283be9e568dc7e91c4ee1ac

Observation 7dd116c8-5efc-4694-ad44-d110294d6535 · outbound

This paper cites Policy gradi- ent methods for reinforcement learning with function approximation,.

Scalable Policy Optimization for Networked Multi-Agent Reinforcement Learning with Continuous State-Action Spaces Policy gradi- ent methods for reinforcement learning with function approximation,

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-01T15:07:45.629412Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T15:07:45.629412Z digest=sha256:14f6aba4633aef21156414ade71e034038807f720ff125964c626f5419e22605

Observation f6d9d64c-2c40-4010-a9b0-ec8248b72543 · outbound

This paper cites Lower bounds for non-convex stochastic optimization,.

Scalable Policy Optimization for Networked Multi-Agent Reinforcement Learning with Continuous State-Action Spaces Lower bounds for non-convex stochastic optimization,

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-01T15:07:45.633092Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T15:07:45.633092Z digest=sha256:00a404f5b8a0b1e5695d2b37acdb5b0e79c38af2ab1dd4d79c96f630d1269f6c

Observation fdf0f9db-840d-4561-9b15-0bcb01c2a9bd · outbound

This paper cites On the theory of policy gradient methods: Optimality, approximation, and distribution shift,.

Scalable Policy Optimization for Networked Multi-Agent Reinforcement Learning with Continuous State-Action Spaces On the theory of policy gradient methods: Optimality, approximation, and distribution shift,

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-01T15:07:45.636986Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T15:07:45.636986Z digest=sha256:676a81fec2fd8b1f2a64f5595293fdfb5b420e2ce0b4b9bc40407c6146302211

Observation 17da105f-f405-4446-bce6-e922b25e8719 · outbound

This paper cites Multi-agent actor-critic for mixed cooperative-competitive environ- ments,.

Scalable Policy Optimization for Networked Multi-Agent Reinforcement Learning with Continuous State-Action Spaces Multi-agent actor-critic for mixed cooperative-competitive environ- ments,

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-01T15:07:45.640729Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T15:07:45.640729Z digest=sha256:13143cc9dc74ae608331a4b7ca20ba19c5775d5716ccf510d6286ad9744c5174

Observation 1ac1c5a9-6649-41a3-8111-6a498f25c506 · outbound

This paper cites Counterfactual multi-agent policy gradients,.

Scalable Policy Optimization for Networked Multi-Agent Reinforcement Learning with Continuous State-Action Spaces Counterfactual multi-agent policy gradients,

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-01T15:07:45.644189Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T15:07:45.644189Z digest=sha256:df12649e431b9f980dbb92006a2e97249b5f06b6e0548e41d0645f9e6aea3b9b

Observation 473a9151-3d76-4f46-a4b3-2dbc7785c245 · outbound

This paper cites Actor-attention-critic for multi-agent reinforcement learning,.

Scalable Policy Optimization for Networked Multi-Agent Reinforcement Learning with Continuous State-Action Spaces Actor-attention-critic for multi-agent reinforcement learning,

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-01T15:07:45.647674Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T15:07:45.647674Z digest=sha256:a1d72b05b4c8646d49ed7e6c13897f3e2964c35d06387c431dac09b38320a03a

Observation b786be04-134f-4043-9584-1e38448ba09a · outbound

This paper cites The surprising effectiveness of ppo in cooperative multi-agent games,.

Scalable Policy Optimization for Networked Multi-Agent Reinforcement Learning with Continuous State-Action Spaces The surprising effectiveness of ppo in cooperative multi-agent games,

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-01T15:07:45.651434Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T15:07:45.651434Z digest=sha256:12be7bb1dafb4e0c5516ca86f346b1f599f2075381a0014bcfaea6065b65efeb

Observation a368e43b-aa00-4bd3-bb82-6b0019cdf913 · outbound

This paper cites Near-optimal distributed linear-quadratic regulator for networked systems,.

Scalable Policy Optimization for Networked Multi-Agent Reinforcement Learning with Continuous State-Action Spaces Near-optimal distributed linear-quadratic regulator for networked systems,

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-01T15:07:45.654811Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T15:07:45.654811Z digest=sha256:9938ccfe8f4995db84200f80402cb7829e93866db8b130065a9d7986942a81d8

Observation 11be9423-6ad9-4224-baf8-250cb9548d5b · outbound

This paper cites Network reconfiguration in distribution systems for loss reduction and load balancing,.

Scalable Policy Optimization for Networked Multi-Agent Reinforcement Learning with Continuous State-Action Spaces Network reconfiguration in distribution systems for loss reduction and load balancing,

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-01T15:07:45.658193Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T15:07:45.658193Z digest=sha256:e59d0ecb306c50a1e42d9f46e929600d3eb60700b96d96aed9c882e3e279526f

Observation 67f10cbf-2aa6-477b-9827-8c23e0c86bdc · outbound

This paper cites The description of a random field by means of conditional probabilities and conditions of its regularity,.

Scalable Policy Optimization for Networked Multi-Agent Reinforcement Learning with Continuous State-Action Spaces The description of a random field by means of conditional probabilities and conditions of its regularity,

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-01T15:07:45.661726Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T15:07:45.661726Z digest=sha256:33928db5262dded96595fdb3aa520ffbe84a149fbccf559784c021b20f8f1211

Observation cb46e0a5-6875-44b3-ba4b-4267b90c209c · outbound

This paper cites Can local particle filters beat the curse of dimensionality?.

Scalable Policy Optimization for Networked Multi-Agent Reinforcement Learning with Continuous State-Action Spaces Can local particle filters beat the curse of dimensionality?

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-01T15:07:45.665084Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T15:07:45.665084Z digest=sha256:5aaa0d9f005c4eee2144ce5d13b675bad84befa1a83bfda07e615e8a41de9253

Observation dde6708c-d21c-42d2-afc5-192f05407b19 · outbound

This paper cites Mini-batch stochastic approx- imation methods for nonconvex stochastic composite optimization,.

Scalable Policy Optimization for Networked Multi-Agent Reinforcement Learning with Continuous State-Action Spaces Mini-batch stochastic approx- imation methods for nonconvex stochastic composite optimization,

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-01T15:07:45.668392Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T15:07:45.668392Z digest=sha256:b95b5382836e5107cd51407154d30ce39f186e56bfccc7d8a35509bf8f35cd68

Observation e0bcf005-760d-4027-9f54-2ae8649d016b · outbound

This paper cites Stochastic first-and zeroth-order methods for nonconvex stochastic programming,.

Scalable Policy Optimization for Networked Multi-Agent Reinforcement Learning with Continuous State-Action Spaces Stochastic first-and zeroth-order methods for nonconvex stochastic programming,

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-01T15:07:45.671697Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T15:07:45.671697Z digest=sha256:ca4bd8e247d30e1b8554c218d95aa998cc0e4becff88d51eb84ff5792e3dfbcd

Pith citing papers

No inbound Pith citation observations are available.