Pith. sign in

Paper Citation Record · LEDGER

Distributed primal-dual algorithm for constrained multi-agent reinforcement learning under coupled policies

As of 17 August 2026, this Paper Citation Record lists 31 of 31 outbound references and 0 inbound Pith citation observations for arXiv:2511.15053.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2511.15053 v3

Coverage vector

measured 31 of 31 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-03T21:36:02.340129Z

measured 31 of 31 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-17T06:30:58.91139+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

31 of 31 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved31
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 202f774e-7096-4356-8161-fbac868d4dc5 · outbound

This paper cites Multiagent deep reinforcement learning for large-scale traffic signal control,.

Distributed primal-dual algorithm for constrained multi-agent reinforcement learning under coupled policies Multiagent deep reinforcement learning for large-scale traffic signal control,

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-03T21:35:58.479380Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T21:35:58.479380Z digest=sha256:7db2648190be9c809ebfdd143a64266da8d661c0cf663e00ff2b88890b89d821

Observation d13c1a4a-3f71-417a-a326-9f89467e8d72 · outbound

This paper cites Large-Scale traffic signal control using a novel multiagent reinforcement learning,.

Distributed primal-dual algorithm for constrained multi-agent reinforcement learning under coupled policies Large-Scale traffic signal control using a novel multiagent reinforcement learning,

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-03T21:35:58.564487Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T21:35:58.564487Z digest=sha256:406139ec6d9746c215eca2fac566a646706b110dd45a95dc81fea9909b905390

Observation c33e2ac0-400b-4c72-92b7-d082f667f0a2 · outbound

This paper cites Applications in traffic signal control: a distributed policy gradient decomposition algorithm,.

Distributed primal-dual algorithm for constrained multi-agent reinforcement learning under coupled policies Applications in traffic signal control: a distributed policy gradient decomposition algorithm,

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-03T21:35:58.675686Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T21:35:58.675686Z digest=sha256:a9772dcc5991739c0321d4f2f8895d1ec1e55f95db835b97636d015d5057cedf

Observation e140e33f-ce4c-4ac2-952d-cf8a50967564 · outbound

This paper cites Deep reinforcement learning for joint channel selection and power control in D2D networks,.

Distributed primal-dual algorithm for constrained multi-agent reinforcement learning under coupled policies Deep reinforcement learning for joint channel selection and power control in D2D networks,

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-03T21:35:58.753478Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T21:35:58.753478Z digest=sha256:7e972d15b4ee211df5a7a94d54f974d772afda77eab34431385d5dbc851611c7

Observation a9b90cb7-8e76-446c-9042-fbcb52c5ffac · outbound

This paper cites Power allocation in multiuser cellular networks: Deep reinforcement learning approaches,.

Distributed primal-dual algorithm for constrained multi-agent reinforcement learning under coupled policies Power allocation in multiuser cellular networks: Deep reinforcement learning approaches,

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-03T21:35:58.872074Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T21:35:58.872074Z digest=sha256:d87255a803bd38c0c73c30b8cc987422b8977ea168cb91e77e4445431cf8f7fd

Observation 6551d4cf-8a2c-49a9-bedf-daf6412588db · outbound

This paper cites Reinforcement learning based recommender systems: a survey,.

Distributed primal-dual algorithm for constrained multi-agent reinforcement learning under coupled policies Reinforcement learning based recommender systems: a survey,

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-03T21:35:59.114391Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T21:35:59.114391Z digest=sha256:d12357736c31325702fdd70462ab26a4a48698a3955a0a4187e9cc46b196f421

Observation f42c7288-b59c-461e-9a3a-cdf6b359ac81 · outbound

This paper cites A survey on reinforcement learning for recommender systems,.

Distributed primal-dual algorithm for constrained multi-agent reinforcement learning under coupled policies A survey on reinforcement learning for recommender systems,

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-03T21:35:59.256275Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T21:35:59.256275Z digest=sha256:97ef1682ba63fecdae254cd52e2e5f1e66681fb789d9f72ec07279b9b743899a

Observation bbdec37a-c4dd-44c0-bd48-4230259fe7d2 · outbound

This paper cites Distributed reinforcement learning algorithm for dynamic economic dispatch with unknown generation cost functions,.

Distributed primal-dual algorithm for constrained multi-agent reinforcement learning under coupled policies Distributed reinforcement learning algorithm for dynamic economic dispatch with unknown generation cost functions,

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-03T21:35:59.447940Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T21:35:59.447940Z digest=sha256:3c1025a69e3a2937f2b3fbb8049a2661265a4305ee41d8d240c659bd35bcddd4

Observation 6b88c68d-1cd0-41ee-ad15-76f428da4d4e · outbound

This paper cites DistributedQ-learning-based online optimization algorithm for unit commitment and dispatch in smart grid,.

Distributed primal-dual algorithm for constrained multi-agent reinforcement learning under coupled policies DistributedQ-learning-based online optimization algorithm for unit commitment and dispatch in smart grid,

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-03T21:35:59.566929Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T21:35:59.566929Z digest=sha256:dfc4b514dfc4d7988b5d13c811269708cfb1f3b60777607533ca88caeb9f1ac1

Observation 9f5a3812-2028-42cd-a30c-b04df87ab78c · outbound

This paper cites Distributed Q-learning algorithm for dynamic resource allocation with unknown objective functions and appli- cation to microgrid,.

Distributed primal-dual algorithm for constrained multi-agent reinforcement learning under coupled policies Distributed Q-learning algorithm for dynamic resource allocation with unknown objective functions and appli- cation to microgrid,

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-03T21:35:59.625353Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T21:35:59.625353Z digest=sha256:6d9306afda82a8834ded5e7a52be2ec0c2cb15905a05d82c5913383950718a12

Observation f0196251-46f1-488b-ab98-d0effe1a17b4 · outbound

This paper cites Safe, Multi-Agent, Reinforcement Learning for Autonomous Driving.

Distributed primal-dual algorithm for constrained multi-agent reinforcement learning under coupled policies Safe, Multi-Agent, Reinforcement Learning for Autonomous Driving

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-03T21:35:59.678018Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T21:35:59.678018Z digest=sha256:da071ebc0f1c909658957794d177cbe027fff0f8048c8cfa3ea7e14a635fc766

Observation c0a90f67-1924-48f4-870c-66dba815b9eb · outbound

This paper cites Con strained reinforcement learning has zero duality gap,.

Distributed primal-dual algorithm for constrained multi-agent reinforcement learning under coupled policies Con strained reinforcement learning has zero duality gap,

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-03T21:35:59.750224Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T21:35:59.750224Z digest=sha256:9e221608882b8bda823298fefd7f7ba5b7be932d6846ef7084920498f40ce019

Observation 76579216-ff66-4fe1-8d2f-1364ca164127 · outbound

This paper cites Safe policies for reinforcement learning via primal-dual meth- ods,.

Distributed primal-dual algorithm for constrained multi-agent reinforcement learning under coupled policies Safe policies for reinforcement learning via primal-dual meth- ods,

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-03T21:35:59.837240Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T21:35:59.837240Z digest=sha256:e943a6cbab661c6d7538cb754630c701507130c59bdffbcf308ff2ebaf356b51

Observation 507e94b4-2379-41f4-885b-6244e5451ab2 · outbound

This paper cites State augmented constrained reinforcement learning: overcoming the limita- tions of learning with rewards,.

Distributed primal-dual algorithm for constrained multi-agent reinforcement learning under coupled policies State augmented constrained reinforcement learning: overcoming the limita- tions of learning with rewards,

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-03T21:36:00.008474Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T21:36:00.008474Z digest=sha256:1fd20e4d5eba89b9a73060277dff8eb966c1da4d3dfbf8904ecb634d51351442

Observation a3cbf48b-7b5b-4b1a-9a0c-5931cbdb46fb · outbound

This paper cites Probabilistic constraint for safety-critical reinforcement learning,.

Distributed primal-dual algorithm for constrained multi-agent reinforcement learning under coupled policies Probabilistic constraint for safety-critical reinforcement learning,

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-03T21:36:00.133026Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T21:36:00.133026Z digest=sha256:b8fcfa41db659ae0b778a9ae00a53bbc26bd2d03b2d5a71cb33709867aca1a9d

Observation a944c460-4235-4e10-9049-9b4bea9313bc · outbound

This paper cites A review of safe reinforcement learning: methods, theories, and applications,.

Distributed primal-dual algorithm for constrained multi-agent reinforcement learning under coupled policies A review of safe reinforcement learning: methods, theories, and applications,

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-03T21:36:00.282863Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T21:36:00.282863Z digest=sha256:4fb3e31684e94e4ee79a77ce3a216617bb9aad5d74936b2cfa53461def2e7c88

Observation 27dbf847-1a01-46c4-991f-c00a01771f26 · outbound

This paper cites Provably efficient safe exploration via primal-dual policy optimization,.

Distributed primal-dual algorithm for constrained multi-agent reinforcement learning under coupled policies Provably efficient safe exploration via primal-dual policy optimization,

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-03T21:36:00.410675Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T21:36:00.410675Z digest=sha256:8e777d5f202defe1d206fdd88c03e833165a8d5be74fa3003de8218ed72f3c80

Observation ef70a023-d5fe-4faf-88a2-7ef0bb92289c · outbound

This paper cites A dual approach to constrained markov decision processes with entropy regularization,.

Distributed primal-dual algorithm for constrained multi-agent reinforcement learning under coupled policies A dual approach to constrained markov decision processes with entropy regularization,

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-03T21:36:00.504081Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T21:36:00.504081Z digest=sha256:b0b30450c08287ab60662be6998f2d908c4f142db70543fbafa63fd208aba77c

Observation a3f90b93-a4ed-4bdc-ab33-7b00d208e04e · outbound

This paper cites Policy-based primal-dual methods for convex constrained markov decision processes,.

Distributed primal-dual algorithm for constrained multi-agent reinforcement learning under coupled policies Policy-based primal-dual methods for convex constrained markov decision processes,

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-03T21:36:00.635532Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T21:36:00.635532Z digest=sha256:8683f8ee232cd19fe7ce4a4e59c8e5100358641f17e5af5b528db1dffec0dc3a

Observation 34992d50-10d7-4009-a6e8-284b92ae6abb · outbound

This paper cites Decentralized policy gradient descent ascent for safe multi-agent reinforcement learning,.

Distributed primal-dual algorithm for constrained multi-agent reinforcement learning under coupled policies Decentralized policy gradient descent ascent for safe multi-agent reinforcement learning,

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-03T21:36:00.879344Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T21:36:00.879344Z digest=sha256:7699ab87819b7d467522a62f8a16d5cb36953585d4e1f4d8c6783c6a091818b3

Observation 4cc5165f-3152-4e34-96ae-8dc5aaf7b2c0 · outbound

This paper cites Scalable primal- dual actor-critic method for safe multi-agent RL with general utilities,.

Distributed primal-dual algorithm for constrained multi-agent reinforcement learning under coupled policies Scalable primal- dual actor-critic method for safe multi-agent RL with general utilities,

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-03T21:36:00.998076Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T21:36:00.998076Z digest=sha256:9e4112c636fcfd5586c96db59c4ddb2b634f30d982ea0f42b39f06ae91067c76

Observation 11f0d457-b2ef-49fc-bac6-1c8e13c3703c · outbound

This paper cites Scalable reinforcement learning of localized policies for multi-agent networked systems,.

Distributed primal-dual algorithm for constrained multi-agent reinforcement learning under coupled policies Scalable reinforcement learning of localized policies for multi-agent networked systems,

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-03T21:36:01.124870Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T21:36:01.124870Z digest=sha256:40b908e906e3cf8690d36167d8becea348e819043367ed9aa158f3955a25af1d

Observation 51f2b191-18ba-4928-9712-df6d1be8f032 · outbound

This paper cites Solving a class of non-convex min-max games using iterative first order methods,.

Distributed primal-dual algorithm for constrained multi-agent reinforcement learning under coupled policies Solving a class of non-convex min-max games using iterative first order methods,

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-03T21:36:01.195030Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T21:36:01.195030Z digest=sha256:35a1a7e2c11567119121c60d12beb32a8fc8d2e43e9c61bffcfa373ff470b4b9

Observation 471f8719-292f-41e7-a22c-eec62f550c55 · outbound

This paper cites Policy gra- dient methods for reinforcement learning with function approximation,.

Distributed primal-dual algorithm for constrained multi-agent reinforcement learning under coupled policies Policy gra- dient methods for reinforcement learning with function approximation,

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-03T21:36:01.293994Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T21:36:01.293994Z digest=sha256:e5aefe8013a22a51fa903b7f923234007246ac0be7a8e7df3dcc1081b0e0d512

Observation 07ed2a0c-8741-4788-bb33-d73021d81a1b · outbound

This paper cites Distributed optimization over time-varying directed graphs,.

Distributed primal-dual algorithm for constrained multi-agent reinforcement learning under coupled policies Distributed optimization over time-varying directed graphs,

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-03T21:36:01.359589Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T21:36:01.359589Z digest=sha256:4ba33d6cb8b5f51e5aeedd1cb48b9484075142cdc7b449ab767d922bbf2d5af7

Observation fd4d33d2-afbe-4190-931c-843e0bbb0e49 · outbound

This paper cites Zeroth-Order Policy Gradient for Reinforcement Learning from Human Feedback without Reward Inference.

Distributed primal-dual algorithm for constrained multi-agent reinforcement learning under coupled policies Zeroth-Order Policy Gradient for Reinforcement Learning from Human Feedback without Reward Inference

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-03T21:36:01.483861Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T21:36:01.483861Z digest=sha256:30b6d9c80aa32d9a5b0f05fd347a02a101a243b9cf1a8816f8838a0e59ada666

Observation 53ab3935-7daa-4efc-a47d-36a7d9d5a556 · outbound

This paper cites ϕ-update: a class of policy update methods with policy convergence guarantee,.

Distributed primal-dual algorithm for constrained multi-agent reinforcement learning under coupled policies ϕ-update: a class of policy update methods with policy convergence guarantee,

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-03T21:36:01.650932Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T21:36:01.650932Z digest=sha256:2410b4aa30136de0cdb63d76e2056c3330c3cf15ea1eb2cc2231a5ae1adcb978

Observation 9d16c18a-1f07-4271-b429-0b7fddc57f83 · outbound

This paper cites Distributed stochastic subgra- dient projection algorithms for convex optimization,.

Distributed primal-dual algorithm for constrained multi-agent reinforcement learning under coupled policies Distributed stochastic subgra- dient projection algorithms for convex optimization,

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-03T21:36:01.788485Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T21:36:01.788485Z digest=sha256:b76e62598205d091d2641133d63aa17f2d305eaa4b6e3035ebcd3a77293fcf13

Observation 2f22a796-18e1-43b1-9b1f-bb53375b693e · outbound

This paper cites Yeh,Real Analysis: Theory of Measure and Integration.

Distributed primal-dual algorithm for constrained multi-agent reinforcement learning under coupled policies Yeh,Real Analysis: Theory of Measure and Integration

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-03T21:36:01.921096Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T21:36:01.921096Z digest=sha256:8438d2506f6b2a8ea3cccfe1301e233cdd2b5941ff0ae093bfe4e72367f4fac6

Observation e6a694b7-2513-4fc3-8b3a-58335de94770 · outbound

This paper cites Global convergence of policy gradient methods to (almost) locally optimal policies,.

Distributed primal-dual algorithm for constrained multi-agent reinforcement learning under coupled policies Global convergence of policy gradient methods to (almost) locally optimal policies,

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-03T21:36:02.149243Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T21:36:02.149243Z digest=sha256:3dd8683a9a44d431b7ebd8a5e6845696e3f0b43ac1087b331da9b82ba5ba7fb9

Observation 6e7dcbf5-7854-4ae8-a71f-9cafcc0e60c2 · outbound

This paper cites Given that the proof procedure is strikingly analogous to that of Case (i), it is omitted for brevity.2 G.

Distributed primal-dual algorithm for constrained multi-agent reinforcement learning under coupled policies Given that the proof procedure is strikingly analogous to that of Case (i), it is omitted for brevity.2 G

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-03T21:36:02.340129Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T21:36:02.340129Z digest=sha256:577afc7805f13207f9e72292479e9fb9e03c24e5aff95a1921d9df7fcf801c0e

Pith citing papers

No inbound Pith citation observations are available.