Pith. sign in

Paper Citation Record · LEDGER

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL

As of 17 August 2026, this Paper Citation Record lists 99 of 99 outbound references and 1 inbound Pith citation observation for arXiv:2504.15425.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2504.15425 v1

Coverage vector

measured 99 of 99 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-16T11:35:49.596258Z

measured 100 of 100 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-16T06:30:59.297886+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-05-13T22:56:33.073457Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-13T22:58:23.930563Z

Reference resolution

99 of 99 outbound references displayed

  • verified exact0
  • verified fuzzy56
  • unresolved43
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 021ed894-e2a8-4fad-af50-233bc9b42bc9 · outbound

This paper cites Constrained policy optimization.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL Constrained policy optimization

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-16T11:35:49.100505Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:35:49.100505Z digest=sha256:d252c26344163f29c6651c5dcba97b745f365e5e74e181642a2762e31ba9e1e1

Observation 3f762109-487a-40e7-a875-1da5a0add65d · outbound

This paper cites Learning transferable cooperative be- havior in multi-agent team.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL Learning transferable cooperative be- havior in multi-agent team

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-16T11:35:49.105943Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:35:49.105943Z digest=sha256:ed027b860301e5e5b13afca74ab3bfdc168d232a204cd5ff7d2cb244d4b6222b

Observation f011d933-34b0-4ec8-956f-40751c62e234 · outbound

This paper cites Constrained Markov decision processes.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL Constrained Markov decision processes

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-16T11:35:49.111335Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:35:49.111335Z digest=sha256:e1f1bb8514499bec59b9bec81c1d734869821009b32474dd10a70cb9c9d854d2

Observation bddfa60f-aab9-4940-a474-151512c7062a · outbound

This paper cites Casadi: a software framework for nonlinear optimization and optimal con- trol.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL Casadi: a software framework for nonlinear optimization and optimal con- trol

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-16T11:35:49.116234Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:35:49.116234Z digest=sha256:3605eb0fc22e5edb2b384873818941b70e05c40fcce1d20092ceff4fc2eb70c4

Observation ba6db136-ebf3-4e72-8c1b-16766ba2e162 · outbound

This paper cites Hamilton-jacobi reachability: A brief overview and recent advances.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL Hamilton-jacobi reachability: A brief overview and recent advances

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-16T11:35:49.121126Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:35:49.121126Z digest=sha256:fc37d29c60475e6d29e9a1b3b18d570b1c17f664578b16c838ad04abfb2102c7

Observation 66e5c626-3936-4ea5-98fb-f527b5901967 · outbound

This paper cites Dynamic programming and optimal control: Volume I, volume 4.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL Dynamic programming and optimal control: Volume I, volume 4

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-16T11:35:49.125942Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:35:49.125942Z digest=sha256:5667976b69d335126857dff161f7b6c706fb7222f4050860ce5922355e5decd5

Observation 647f9e09-641c-4aa1-b5cb-6078b0a34493 · outbound

This paper cites Synthesis of minimum-cost shields for multi-agent systems.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL Synthesis of minimum-cost shields for multi-agent systems

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-16T11:35:49.131262Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:35:49.131262Z digest=sha256:9100e606d093db257cf5afec31f730a2236186098c1ed9ad23d5addaa04c3d22

Observation 01d8c8e2-5ed1-4d53-b4da-b702c9846195 · outbound

This paper cites An actor-critic algorithm for constrained markov decision processes.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL An actor-critic algorithm for constrained markov decision processes

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-16T11:35:49.136145Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:35:49.136145Z digest=sha256:e635faa268484a275ab4a81d9c05c99f9332d768c0c3624e5e127456982a301a

Observation 387cca71-61aa-4619-9945-16a5dd173f73 · outbound

This paper cites Stochastic Approximation: A Dynamical Systems Viewpoint, volume 48.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL Stochastic Approximation: A Dynamical Systems Viewpoint, volume 48

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-16T11:35:49.141464Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:35:49.141464Z digest=sha256:0057551c45291a9893430b5496c0083c0fd094ec0593a41f76e9c472ae7abe19

Observation f2eede24-1a56-4d49-8870-c9075d9289b3 · outbound

This paper cites Convex optimization.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL Convex optimization

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-16T11:35:49.146263Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:35:49.146263Z digest=sha256:1b819479d1be03b4bd6834a4be33356ceee9afee32b76e2969e3edd02d1b6c00

Observation b3dae811-fdde-4419-b16c-879328c11f8f · outbound

This paper cites Safe Multi-Agent Reinforcement Learning through Decentralized Multiple Control Barrier Functions.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL Safe Multi-Agent Reinforcement Learning through Decentralized Multiple Control Barrier Functions

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-16T11:35:49.151496Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:35:49.151496Z digest=sha256:e3ef3900d650bcbbe94c3030de3f23b110cea3f966c8dbd46879bdf1aa48d3f5

Observation 47be9f9e-05d5-44dd-8c33-ac89b9305b93 · outbound

This paper cites A new hybrid quadratic/bisection algorithm for finding the zero of a nonlinear function without using derivatives.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL A new hybrid quadratic/bisection algorithm for finding the zero of a nonlinear function without using derivatives

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-16T11:35:49.156876Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:35:49.156876Z digest=sha256:67011cd3ba72930a960909c0332dbd5ef4dc65bff0f7fb7696ad026d387324eb

Observation 1182b6e9-eee5-4d80-9d2c-bf4c05b5cc3c · outbound

This paper cites Socially aware motion planning with deep re- inforcement learning.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL Socially aware motion planning with deep re- inforcement learning

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-16T11:35:49.161771Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:35:49.161771Z digest=sha256:b326335c504771c6a2151f547be1446870469e58e9a592f25756e47273cfc72a

Observation 4924cc63-5136-465e-8dee-b868114ab8ec · outbound

This paper cites Decentralized non-communicating multiagent col- lision avoidance with deep reinforcement learning.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL Decentralized non-communicating multiagent col- lision avoidance with deep reinforcement learning

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-16T11:35:49.166199Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:35:49.166199Z digest=sha256:c7561c252c62fb6bc9d277a55eac59cca8bc7c1beed42c2471499accb44f7130

Observation 2986275c-4cda-40fe-b0ac-f6f90ac66d74 · outbound

This paper cites On the duality gap of constrained cooperative multi-agent reinforcement learning.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL On the duality gap of constrained cooperative multi-agent reinforcement learning

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-16T11:35:49.170705Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:35:49.170705Z digest=sha256:3a679b69fab7f4cc0d5b4fb219ca66cf198d86ef887d061523ab3e2b631a1fd6

Observation dfd586fa-ea56-4f86-b855-92705b94571f · outbound

This paper cites Computational aspects of distributed optimization in model predictive control.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL Computational aspects of distributed optimization in model predictive control

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-16T11:35:49.175559Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:35:49.175559Z digest=sha256:47729d23e4d73fe0dfe9f72d988737d080f5413a2d1a8358fa079d418b1f131c

Observation e1c51361-7511-40ce-b559-ece92db3f9bd · outbound

This paper cites De- tecting, localizing, and tracking an unknown number of moving targets using a team of mobile robots.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL De- tecting, localizing, and tracking an unknown number of moving targets using a team of mobile robots

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-16T11:35:49.180383Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:35:49.180383Z digest=sha256:c2716b5e744a7bf23752390fcb66ecc5a2a2c87596041705d89b6c3c682064f6

Observation 357a9e3c-ec56-4522-860a-ca97b7967865 · outbound

This paper cites Provably efficient gener- alized lagrangian policy optimization for safe multi-agent reinforcement learning.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL Provably efficient gener- alized lagrangian policy optimization for safe multi-agent reinforcement learning

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-16T11:35:49.185205Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:35:49.185205Z digest=sha256:85cd3e034fbe88e82454127ab68211a113a9f2d25df1fd1b64c13fc4a5219703

Observation 81d691fe-8670-41e6-b37e-f155dddd9993 · outbound

This paper cites Safe Multi-Agent Reinforcement Learning via Shielding.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL Safe Multi-Agent Reinforcement Learning via Shielding

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-16T11:35:49.190136Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:35:49.190136Z digest=sha256:dfc29ead2c7974a3e13ca1f8b9c8466c56ec4b5d1fed443de3881a474e295cb1

Observation e0024367-a868-409e-8988-6a3c1dbe99a2 · outbound

This paper cites Safe multi- agent reinforcement learning via shielding.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL Safe multi- agent reinforcement learning via shielding

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-16T11:35:49.194900Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:35:49.194900Z digest=sha256:edd9cad656325432509b9faa49516a477acbe9d451b1243d00b717e3bdd02524

Observation 5101e1bb-d11f-46c4-b075-44e424cad64b · outbound

This paper cites Mo- tion planning among dynamic, decision-making agents with deep reinforcement learning.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL Mo- tion planning among dynamic, decision-making agents with deep reinforcement learning

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-16T11:35:49.199462Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:35:49.199462Z digest=sha256:91e66e2993144a6c1d0bba8d4e88210e41a98fc498a016f558796b74cddddbe6

Observation a2625dc9-4098-48a1-a000-2456fb48fdf1 · outbound

This paper cites A distributed model predictive control strategy for constrained multi- agent systems: The uncertain target capturing scenario.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL A distributed model predictive control strategy for constrained multi- agent systems: The uncertain target capturing scenario

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:35:50.960217Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-16T11:35:49.203864Z digest=sha256:d36fe6a9c4ecd1cbf10bd6cf9afce9be2b3bc71e1fa3443d703b2ad13bb5a4b8

Observation a37f3fec-48f4-496b-9e91-42bd4d4dca69 · outbound

This paper cites Counterfactual multi-agent policy gradients.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL Counterfactual multi-agent policy gradients

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:35:50.944198Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-16T11:35:49.208585Z digest=sha256:bfcd5425372de56227a518773713119ec5476a5b1a72f1d70992fe5054b1618d

Observation a36e666d-7a33-4222-aba9-4073d813675b · outbound

This paper cites Iterative reachability estimation for safe reinforcement learning.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL Iterative reachability estimation for safe reinforcement learning

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-16T11:35:49.213377Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:35:49.213377Z digest=sha256:d9097eaa0800e63826f1ccf4a0cfaf1769db1271707117c6ed8433cad58c1691

Observation 3ca00963-31c1-42f3-99d1-b440454e11c5 · outbound

This paper cites Learning safe control for multi- robot systems: Methods, verification, and open chal- lenges.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL Learning safe control for multi- robot systems: Methods, verification, and open chal- lenges

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:35:50.916330Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-16T11:35:49.218177Z digest=sha256:6a6f3a3e77acae2b795b7dd7e7fd465b55d09563d9360147c3cf01be700f6349

Observation 2e19ab2a-784e-41e9-9806-908398d7a7fc · outbound

This paper cites A reinforce- ment learning framework for vehicular network routing under peak and average constraints.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL A reinforce- ment learning framework for vehicular network routing under peak and average constraints

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:35:50.899242Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-16T11:35:49.222823Z digest=sha256:5e78d9e21270b1542194393b752080e3b9f176bf91d958fc2d809302bbde2f36

Observation 7e679174-932b-4bb9-9733-7274dc2c81d5 · outbound

This paper cites Crazyflie 2.0 quadrotor as a platform for research and education in robotics and control engineering.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL Crazyflie 2.0 quadrotor as a platform for research and education in robotics and control engineering

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:35:50.882587Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-16T11:35:49.228778Z digest=sha256:43ae152a3a129224d31949bca636c86b6ba8bfd5deb3d858187a98688706f2e1

Observation 6f6e2a00-a966-4689-a670-3e1230b068cd · outbound

This paper cites Snopt: An sqp algorithm for large-scale constrained optimization.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL Snopt: An sqp algorithm for large-scale constrained optimization

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-16T11:35:49.234123Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:35:49.234123Z digest=sha256:95ca8a6e9716484691dd1d88ee3b7b1e901140a3edbd1dce4fcda1466c981531

Observation c7c38f09-8592-41a2-b2e7-97f7db9891a4 · outbound

This paper cites Nonlinear model predictive control: theory and algorithms.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL Nonlinear model predictive control: theory and algorithms

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:35:50.853203Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-16T11:35:49.239025Z digest=sha256:1dbfd6974b60f8a703698d8e4c1dc3732a408b102d32c86813390128223d63fa

Observation cd1c5b7b-32eb-4073-a64f-0fea3ec9c907 · outbound

This paper cites Multi-Agent Constrained Policy Optimisation.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL Multi-Agent Constrained Policy Optimisation

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-16T11:35:49.243705Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:35:49.243705Z digest=sha256:e183acbf46b8e4d1438bc69ec4b520c1b17f486921a3df547a141045ebf13abf

Observation 75bcca73-1d38-4d39-b5ae-6474eddf8218 · outbound

This paper cites A Review of Safe Reinforcement Learning: Methods, Theory and Applications.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL A Review of Safe Reinforcement Learning: Methods, Theory and Applications

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-16T11:35:49.248824Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:35:49.248824Z digest=sha256:d4e9bacff836c54bcb6f5085a65bd5e3b02b8cb4ff9a2ee7986f2b2ae9efc19c

Observation df65dcad-5e97-4d33-84b7-8389674db06d · outbound

This paper cites Safe multi-agent reinforcement learning for multi-robot control.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL Safe multi-agent reinforcement learning for multi-robot control

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:35:50.834594Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-16T11:35:49.253790Z digest=sha256:b2d4d25d778d32850913ad0aa5451f38b21471a784459a4083ee6e74ddf116cb

Observation 7fe7dc21-ca19-41c5-a1c5-f1443941b0a2 · outbound

This paper cites Coordinated reinforcement learning.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL Coordinated reinforcement learning

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:35:50.816795Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-16T11:35:49.258380Z digest=sha256:a1370914b9af689ced3308818385550775aa23138765c469d3e0bd22ca809711

Observation bce72a09-d609-4b10-9671-4cce7b9d10ed · outbound

This paper cites Deep recurrent q-learning for partially observable mdps.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL Deep recurrent q-learning for partially observable mdps

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:35:50.800028Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-16T11:35:49.263244Z digest=sha256:9149906eea3c84f2486172c5d436710f720d27e017f1644ff29df0fc67f502f3

Observation efc781f4-8330-4f9e-8ab0-b0fa59a789d6 · outbound

This paper cites Autocost: Evolving intrinsic cost for zero-violation reinforcement learning.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL Autocost: Evolving intrinsic cost for zero-violation reinforcement learning

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:35:50.782922Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-16T11:35:49.269034Z digest=sha256:f447d547aeefc0fa54bd9676d6e865057b5e7a9741753c86f924c8ec89e3c3fa

Observation 60b7bea2-7cf5-4183-ab78-6e180f5141a5 · outbound

This paper cites Safedreamer: Safe reinforcement learning with world models.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL Safedreamer: Safe reinforcement learning with world models

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:35:50.765390Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-16T11:35:49.273968Z digest=sha256:43e4c78cac71a3938ae236673a6ea827d815fd4ea5a4e1a199ab541afb1adf3d

Observation bcef9d51-07e5-449c-8798-89fdffd4bdb4 · outbound

This paper cites Distributed optimization in multi-agent robotics for industry 4.0 warehouses.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL Distributed optimization in multi-agent robotics for industry 4.0 warehouses

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:35:50.747003Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-16T11:35:49.280957Z digest=sha256:da7fc43b0c92630b9bc56fc06cf1b0af49e694acfada0764d8cb99c63c0df146

Observation ea097793-3e3b-4905-be52-a3a255c5c929 · outbound

This paper cites Cmix: Deep multi- agent reinforcement learning with peak and average con- straints.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL Cmix: Deep multi- agent reinforcement learning with peak and average con- straints

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:35:50.728160Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-16T11:35:49.286188Z digest=sha256:c2c3fc6555e84dc86c18606015fdf889e7393fd48227634fca76cbc82e04e078

Observation 1406ba9c-9992-4145-9e9c-3710efa7d9fc · outbound

This paper cites Towards optimally decentralized multi-robot collision avoidance via deep reinforcement learning.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL Towards optimally decentralized multi-robot collision avoidance via deep reinforcement learning

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:35:50.705889Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-16T11:35:49.290715Z digest=sha256:e1c62e2719172692e4b8ab337041943099b1336780957844bdd31a717a53ac44

Observation 08f7fcea-19d9-480d-a403-5c4ed1d81e61 · outbound

This paper cites Multi-agent actor- critic for mixed cooperative-competitive environments.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL Multi-agent actor- critic for mixed cooperative-competitive environments

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:35:50.687627Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-16T11:35:49.295450Z digest=sha256:d130ffdd60fb4e1b1f8d367d654dbf458680ad762de7ac49004cd04d4584df6c

Observation 1ec2c0d6-124d-4e49-8d1c-7e4796f40a9c · outbound

This paper cites Decentralized policy gradient descent ascent for safe multi-agent reinforcement learning.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL Decentralized policy gradient descent ascent for safe multi-agent reinforcement learning

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:35:50.670639Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-16T11:35:49.300993Z digest=sha256:c9e6d5f7b544fbeed6598d028dd70e9f7d97b198f69fdaf2b0f131911d638ecf

Observation 37f2e7e6-78ba-46cb-8ac7-34d294efe1ca · outbound

This paper cites Trajectory generation for multiagent point-to-point transitions via distributed model predictive control.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL Trajectory generation for multiagent point-to-point transitions via distributed model predictive control

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:35:50.653222Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-16T11:35:49.306462Z digest=sha256:aa48093ed19832e6e2b873cadb47bec798d9987eb516329f9676b010eaf505b3

Observation 94b03caa-0cb1-4e78-8c51-3b1b3280dc16 · outbound

This paper cites Online trajectory generation with distributed model predictive control for multi-robot motion planning.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL Online trajectory generation with distributed model predictive control for multi-robot motion planning

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:35:50.634565Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-16T11:35:49.311335Z digest=sha256:251282958fbf9dd4a1068e9ebd934f9d44f2d6f4034b801d7a4509c857684dfb

Observation 3d62944b-59f5-43d1-a914-1388f859acdf · outbound

This paper cites On reachability and minimum cost optimal control.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL On reachability and minimum cost optimal control

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-16T11:35:49.316420Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:35:49.316420Z digest=sha256:a5be2abcd4fb366f7e0157298e32c5d0976511fc5b27a8c909e04f6fd05af856

Observation a67092fb-f45b-4168-8ce2-1c802594f8d9 · outbound

This paper cites Lifelong Multi-Agent Path Finding for Online Pickup and Delivery Tasks.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL Lifelong Multi-Agent Path Finding for Online Pickup and Delivery Tasks

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-16T11:35:49.321247Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:35:49.321247Z digest=sha256:6012279fc0cece438fc802e3b1f5091ed9fd850172e1f09abe7baee6de6abe0b

Observation 93311ed5-1199-4c60-9bf5-029f404aa922 · outbound

This paper cites Hamilton–jacobi formulation for reach–avoid differential games.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL Hamilton–jacobi formulation for reach–avoid differential games

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:35:50.605556Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-16T11:35:49.326900Z digest=sha256:a411c6633ee08029e78e4463865eb239df35d3b9490abdb94985a65b838f2f4c

Observation bfd02cfa-a22b-4953-be02-47eba0a8793a · outbound

This paper cites Safe value functions.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL Safe value functions

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:35:50.587653Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-16T11:35:49.331833Z digest=sha256:6e2e59d4eaafcda73bbd2add71c30161c1b5acaa6addba5a7436aff2da645fde

Observation bc854208-b398-4db6-beec-579021b599d8 · outbound

This paper cites Shield decentralization for safe multi-agent reinforce- ment learning.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL Shield decentralization for safe multi-agent reinforce- ment learning

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:35:50.570150Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-16T11:35:49.337215Z digest=sha256:4cfb6a0379b5cc8778fdc40bd24f65b8babfe79167b61acabbb3b5cea6c8e6a5

Observation 140a3280-0ba2-4594-8868-3c5926e56dfd · outbound

This paper cites A time-dependent hamilton-jacobi formulation of reachable sets for continuous dynamic games.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL A time-dependent hamilton-jacobi formulation of reachable sets for continuous dynamic games

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-16T11:35:49.342067Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:35:49.342067Z digest=sha256:24bdb714dedce8a33de6a041c1de0e0762c7b909a5064dc0fb591f33a1378e5e

Observation e0fec4a7-ff8b-4d82-a8bc-93d754141da9 · outbound

This paper cites Distributed model predictive safety certification for learning-based control.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL Distributed model predictive safety certification for learning-based control

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:35:50.540428Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-16T11:35:49.346763Z digest=sha256:30d0713dc55ec6aa2b58a319b959737a540324ff9c889fc8e2b4e8dcd616440e

Observation 2f02698a-e874-44db-974c-e81feffb43b8 · outbound

This paper cites Scalable multi-agent reinforcement learning through intelligent information aggregation.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL Scalable multi-agent reinforcement learning through intelligent information aggregation

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:35:50.523861Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-16T11:35:49.351927Z digest=sha256:34957cbd8a2d3088342d749b2b1d30e03b2c51db36f9a6062a087a8add730c50

Observation fd4865da-e1d0-4b21-a8d3-347e590770d3 · outbound

This paper cites Distributed optimization for control.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL Distributed optimization for control

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:35:50.506329Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-16T11:35:49.356555Z digest=sha256:7c128a6fa410bbcb50578496ed76e29784bdbbfe6c838ae7bcf2f9d9ad131c05

Observation a9691ad1-2670-49cf-b45e-2c22cdba0e9c · outbound

This paper cites Numerical opti- mization.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL Numerical opti- mization

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-16T11:35:49.361981Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:35:49.361981Z digest=sha256:e527ae701a2b43432259db6d482b5ffa0e1432a9f342039d961404c2f7824230

Observation 13192686-87fe-4a3b-9a5c-6f1a74d3a50f · outbound

This paper cites Facmac: Factored multi- agent centralised policy gradients.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL Facmac: Factored multi- agent centralised policy gradients

Reference 54

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:35:50.479369Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-16T11:35:49.366973Z digest=sha256:2527ff55756a49a0d81ec0ce491c1367a89777c1ec70588b978f1cc3bcc04f28

Observation 21340052-cdf4-4616-b3b5-61c0855e9881 · outbound

This paper cites Decentralized Safe Multi-agent Stochastic Optimal Control using Deep FBSDEs and ADMM.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL Decentralized Safe Multi-agent Stochastic Optimal Control using Deep FBSDEs and ADMM

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-16T11:35:49.372051Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:35:49.372051Z digest=sha256:97d5c817dc1777e159ee19bb662c0f17f29477d9a360123acbae1be0cad875b6

Observation aabf5350-5c2b-4d98-8168-1bd78439b3fb · outbound

This paper cites Learning safe multi-agent control with decentralized neural barrier certificates.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL Learning safe multi-agent control with decentralized neural barrier certificates

Reference 56

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:35:50.462917Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-16T11:35:49.377313Z digest=sha256:3b8f899a84bad61c4a2665e01209ce4fd36d9de6e43ccfa3b299f749af7ae5b4

Observation f2d235d6-6442-4c49-b465-007add8b0b91 · outbound

This paper cites Weighted qmix: Expanding monotonic value function factorisation for deep multi-agent reinforcement learning.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL Weighted qmix: Expanding monotonic value function factorisation for deep multi-agent reinforcement learning

Reference 57

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:35:50.446467Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-16T11:35:49.382001Z digest=sha256:0e1c23eaf8e2b6f1e739ef1e6346c11272ef9566c85d36fe5186f005de4d36ed

Observation 9146e6a4-38c0-47b9-b853-b50494f6c11c · outbound

This paper cites Monotonic value function factorisation for deep multi-agent reinforcement learning.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL Monotonic value function factorisation for deep multi-agent reinforcement learning

Reference 58

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:35:50.430474Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-16T11:35:49.386675Z digest=sha256:18b986ac0614e3b22e77f50f013fc463b92dcbd2b92c1ebf1685a755efa45a2d

Observation bd3fb1fe-ecc5-4994-8921-57747eb556e1 · outbound

This paper cites A stochastic approx- imation method.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL A stochastic approx- imation method

Reference 59

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:35:50.414429Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-16T11:35:49.391565Z digest=sha256:d7b4473e86e32e81719bf8b369b7f4c7f7ef87c6a8ed1941d32b9fa5655a97e0

Observation 3c9c3079-afe6-4d49-845b-79ebf0a70db0 · outbound

This paper cites Con- strained markov decision processes via backward value functions.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL Con- strained markov decision processes via backward value functions

Reference 60

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:35:50.398173Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-16T11:35:49.396466Z digest=sha256:aebeb5d2c44f5b42d7cf9e0718c79ebb65b5733e00582cc46d4daf86e9a0b40a

Observation d6075347-e324-4dba-8847-a101ca763e47 · outbound

This paper cites Trust region policy optimiza- tion.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL Trust region policy optimiza- tion

Reference 61

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:35:50.381182Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-16T11:35:49.401269Z digest=sha256:9f87675ee6cef32bd8472ef3ec90633911e235ef13087fa1db15c177897e5bcc

Observation 3b10c4ad-aecd-42e6-93e6-5b509e001dc4 · outbound

This paper cites High-Dimensional Continuous Control Using Generalized Advantage Estimation.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL High-Dimensional Continuous Control Using Generalized Advantage Estimation

Reference 62

Resolution
unresolved
no resolver link, observed 2026-08-16T11:35:49.405908Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:35:49.405908Z digest=sha256:f6365b7fe3254b50b05f88fd23f2204420fde090f77eddb09b2d084d6f438eda

Observation fee63235-7e9a-4da2-a553-945bc6e6ac5b · outbound

This paper cites Proximal Policy Optimization Algorithms.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL Proximal Policy Optimization Algorithms

Reference 63

Resolution
unresolved
no resolver link, observed 2026-08-16T11:35:49.411092Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:35:49.411092Z digest=sha256:61509e8010eb60e4d1d38cad97b9bbe3b8422d52af646490a83eea6f5909854b

Observation 3adb61f4-599e-4493-8fef-ffab520f8d3c · outbound

This paper cites Multi-agent motion planning for dense and dynamic environments via deep reinforcement learning.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL Multi-agent motion planning for dense and dynamic environments via deep reinforcement learning

Reference 64

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:35:50.363450Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-16T11:35:49.415932Z digest=sha256:e566ecb33d718b28b8613f849e825cfe3302941843cd680764a89ac220a5705d

Observation 41fad7e6-d630-4b0c-ba21-c3394c99f9dd · outbound

This paper cites Masked Label Prediction: Unified Message Passing Model for Semi-Supervised Classification.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL Masked Label Prediction: Unified Message Passing Model for Semi-Supervised Classification

Reference 65

Resolution
unresolved
no resolver link, observed 2026-08-16T11:35:49.420858Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:35:49.420858Z digest=sha256:c7a014dc86d5de1bcf4bebafd0c5ba349032f7c38ccbcbe378deb860d46bbee4

Observation 1309f659-a663-49a7-8ca9-119b5083d7bb · outbound

This paper cites Solving stabilize-avoid optimal control via epigraph form and deep reinforce- ment learning.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL Solving stabilize-avoid optimal control via epigraph form and deep reinforce- ment learning

Reference 66

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:35:50.347386Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-16T11:35:49.426325Z digest=sha256:f72b51a0a554e4328288d88298d083b4360dd746f4cf2483714377101fe9de8f

Observation 29aebdd6-f22a-4bed-ab12-78370cd16582 · outbound

This paper cites Solving minimum-cost reach avoid using reinforcement learning.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL Solving minimum-cost reach avoid using reinforcement learning

Reference 67

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:35:50.330772Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-16T11:35:49.431478Z digest=sha256:7835b90d4fd97422a8ef96e2dc28b86eef1df664c32c5457a6da616177be31b8

Observation bf9028ad-636b-40e7-baee-5aa78a81ff94 · outbound

This paper cites Predictive control of aerial swarms in cluttered environ- ments.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL Predictive control of aerial swarms in cluttered environ- ments

Reference 68

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:35:50.311938Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-16T11:35:49.436382Z digest=sha256:7db3defefc569770e9e153f3246ff45206c71041ae6cc67e242c155e6502bc70

Observation 1926dfd9-129d-44d3-9ebb-4915cc2c9790 · outbound

This paper cites Value-Decomposition Networks For Cooperative Multi-Agent Learning.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL Value-Decomposition Networks For Cooperative Multi-Agent Learning

Reference 69

Resolution
unresolved
no resolver link, observed 2026-08-16T11:35:49.441863Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:35:49.441863Z digest=sha256:5ad4efcd0692cce1484f10c2911e7515f972e7f8492c7fd41e16d38e69a5a7de

Observation 54313299-44bc-4c9e-ba05-01afdf70c610 · outbound

This paper cites Mankowitz, and Shie Mannor.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL Mankowitz, and Shie Mannor

Reference 70

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:35:50.293658Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-16T11:35:49.447051Z digest=sha256:ce678cfd4b6fac8ab6fd32803ac37d85f907a8424a01e5da6a6179327b990329

Observation 1522d750-8cc9-46c0-92a8-36bb1816632f · outbound

This paper cites A game theoretic approach to controller design for hybrid systems.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL A game theoretic approach to controller design for hybrid systems

Reference 71

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:35:50.276454Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-16T11:35:49.452061Z digest=sha256:87b7f5f126c938bd387a2b4a410b92bb946f480c9c4bd65e412c09f58cf3e252

Observation fcfdc6fa-dcc9-44db-82af-7ef417761499 · outbound

This paper cites Decentralized multi-agent planning using model predictive control and time-aware safe corridors.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL Decentralized multi-agent planning using model predictive control and time-aware safe corridors

Reference 72

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:35:50.260275Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-16T11:35:49.456929Z digest=sha256:6ffab3c83ff691f11ba1d72cd35f0d03ea623b1920629f8ac3ed3f4c086ecc8d

Observation f2d69c81-cba1-4274-9b0a-19aa5399a1d1 · outbound

This paper cites Initial guess generation for aircraft landing trajec- tory optimization.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL Initial guess generation for aircraft landing trajec- tory optimization

Reference 73

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:35:50.241930Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-16T11:35:49.461859Z digest=sha256:0b13249d52b4c31828d6568613e2d455f6ef65668066fb87c5932e6fe5f52394

Observation 85c5ed1a-92ab-48d1-ad9b-02b126e0355f · outbound

This paper cites QPLEX: Duplex Dueling Multi-Agent Q-Learning.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL QPLEX: Duplex Dueling Multi-Agent Q-Learning

Reference 74

Resolution
unresolved
no resolver link, observed 2026-08-16T11:35:49.466812Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:35:49.466812Z digest=sha256:237ca22a567de8b07c66da22535e276e16999f70b4fa7157435e1535692a3dc5

Observation 587b983f-2f82-460d-b4f8-21478fb37347 · outbound

This paper cites A synthesis approach of distributed model predictive control for homogeneous multi-agent system with collision avoidance.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL A synthesis approach of distributed model predictive control for homogeneous multi-agent system with collision avoidance

Reference 75

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:35:50.225310Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-16T11:35:49.472565Z digest=sha256:ef885fd051f5a82020591cfb7625510d568ed3cdb5f0edfd1eb98da5c265002f

Observation 04916c6e-b7d6-41bc-be07-e2a7093bf891 · outbound

This paper cites Multi-agent deep reinforcement learning for urban traffic light control in vehicular networks.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL Multi-agent deep reinforcement learning for urban traffic light control in vehicular networks

Reference 76

Resolution
unresolved
no resolver link, observed 2026-08-16T11:35:49.478723Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:35:49.478723Z digest=sha256:11756db00305e7bd42bc7e235caeb7d4f8a05bfd8ebeb4d70952526008782643

Observation abafd294-d515-4529-a413-1c0fd57fab81 · outbound

This paper cites Model-based Dynamic Shielding for Safe and Efficient Multi-Agent Reinforcement Learning.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL Model-based Dynamic Shielding for Safe and Efficient Multi-Agent Reinforcement Learning

Reference 77

Resolution
unresolved
no resolver link, observed 2026-08-16T11:35:49.483573Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:35:49.483573Z digest=sha256:7240b7338aef9d4ebf7de0f436a35bbcd95c115454855e4615b740472e7cb1b6

Observation 1f6914fa-a1f2-4d92-8206-66da2e5e14e4 · outbound

This paper cites Crpo: A new approach for safe reinforcement learning with convergence guarantee.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL Crpo: A new approach for safe reinforcement learning with convergence guarantee

Reference 78

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:35:50.196643Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-16T11:35:49.488687Z digest=sha256:61142acb647f90386bcac3d4285a24023cbde6e8e6750ed8dcaf2d6fc0bab82b

Observation 71321466-c55e-4511-a348-002b28d81941 · outbound

This paper cites Qatten: A General Framework for Cooperative Multiagent Reinforcement Learning.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL Qatten: A General Framework for Cooperative Multiagent Reinforcement Learning

Reference 79

Resolution
unresolved
no resolver link, observed 2026-08-16T11:35:49.493384Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:35:49.493384Z digest=sha256:1f87c26864d034159dd29cc11b178e439e79280b720d74221e4c58bdb09ef555

Observation 69237a14-51b7-4de9-a903-c724df5cc2e1 · outbound

This paper cites The surprising effectiveness of ppo in cooperative multi-agent games.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL The surprising effectiveness of ppo in cooperative multi-agent games

Reference 80

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:35:50.178088Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-16T11:35:49.499001Z digest=sha256:7fcde75fda5b2880c1690a9e88554eac7c8d746771010a8c65821d1f8210cd39

Observation 9095540a-b7ea-4b7b-9411-59b023a5f2e6 · outbound

This paper cites Reachability constrained reinforcement learning.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL Reachability constrained reinforcement learning

Reference 81

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:35:50.161635Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-16T11:35:49.504513Z digest=sha256:8e12d841280f670947beffd91c125a75806a5147a9b549eb9013821f86474470

Observation b2c0f567-268e-4b73-9e38-2d1f76ccf2b8 · outbound

This paper cites Safe reinforcement learning using robust mpc.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL Safe reinforcement learning using robust mpc

Reference 82

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:35:50.143568Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-16T11:35:49.509472Z digest=sha256:31442748ac6db07880abece0f470311783b2f1f677bb80cd392477f6e2ad81f5

Observation 73060d35-0d38-480d-bfe0-fee0857b7def · outbound

This paper cites Fully decentralized multi-agent re- inforcement learning with networked agents.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL Fully decentralized multi-agent re- inforcement learning with networked agents

Reference 83

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:35:50.124360Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-16T11:35:49.514827Z digest=sha256:bd411a8e8cac1d1155f0196f810772af80729c6e8294d6fb676eea508307a91d

Observation ffe68379-54e4-4d0b-a739-3bd8773eda9e · outbound

This paper cites Multi- agent reinforcement learning: A selective overview of theories and algorithms.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL Multi- agent reinforcement learning: A selective overview of theories and algorithms

Reference 84

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:35:50.105282Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-16T11:35:49.520678Z digest=sha256:90714a5cbcd9c8e5ac480ec8499ea8d72eeb35bf7066ce8acb774e2bb586e468

Observation 6eb47d82-ce9b-450e-a34e-ae2cb5161338 · outbound

This paper cites Neu- ral graph control barrier functions guided distributed collision-avoidance multi-agent control.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL Neu- ral graph control barrier functions guided distributed collision-avoidance multi-agent control

Reference 85

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:35:50.088952Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-16T11:35:49.525301Z digest=sha256:29813f422a2cde6e71d51fa8796ae59f98ca7b273f88370fde8686d448c7192b

Observation cc55ae64-209c-4442-a094-ae8e0610b2d1 · outbound

This paper cites Discrete GCBF proximal policy optimization for multi-agent safe optimal control.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL Discrete GCBF proximal policy optimization for multi-agent safe optimal control

Reference 86

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:35:50.072861Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-16T11:35:49.530170Z digest=sha256:fb2b8851c0e5a36b6f4fbefc746819234b24676326855735ce0030131436c07b

Observation 56eb93fb-3c0f-401d-b7cd-674dd7a7498e · outbound

This paper cites GCBF+: A neural graph control barrier function framework for distributed safe multiagent control.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL GCBF+: A neural graph control barrier function framework for distributed safe multiagent control

Reference 87

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:35:50.055969Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-16T11:35:49.534940Z digest=sha256:3798b8339f25a7fd8c3c3520cea1c95ce99b0f6fb730747ea90e0430d84f69f2

Observation 89e57848-5be3-4969-8f1a-b94fca74a625 · outbound

This paper cites MAMPS: Safe Multi-Agent Reinforcement Learning via Model Predictive Shielding.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL MAMPS: Safe Multi-Agent Reinforcement Learning via Model Predictive Shielding

Reference 88

Resolution
unresolved
no resolver link, observed 2026-08-16T11:35:49.539360Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:35:49.539360Z digest=sha256:a47d717cc223b126adaa1355add0e58112df1272c14aec418b2e897c233885c6

Observation cc1d6b07-680b-41d2-9377-4f86753bcc21 · outbound

This paper cites Model-free safe control for zero-violation reinforcement learning.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL Model-free safe control for zero-violation reinforcement learning

Reference 89

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:35:50.040156Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-16T11:35:49.544295Z digest=sha256:3927590dff227c32fab6ec802fa4f8fd6c6b5f9a1f4ea2d6d3b2128fda493f81

Observation 02e63a0a-c108-48df-95a8-786b32f80c81 · outbound

This paper cites Multi-agent first order con- strained optimization in policy space.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL Multi-agent first order con- strained optimization in policy space

Reference 90

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:35:50.023762Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-16T11:35:49.549368Z digest=sha256:6d4850951ade46fdad3b72a5eedad86da1e687207e5a42933d445532bf950944

Observation c23fb004-13ae-4b79-bebd-59a7e5ab4f6f · outbound

This paper cites Fast, on-line collision avoidance for dynamic vehicles using buffered voronoi cells.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL Fast, on-line collision avoidance for dynamic vehicles using buffered voronoi cells

Reference 91

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:35:50.006851Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-16T11:35:49.555561Z digest=sha256:e7e39e130a38ca3be6b74002a963e944f47c0e2cb4b35547e5dcceac0d2c061d

Observation d8b0142c-0fc5-41ba-b1f1-3b1efdd930f5 · outbound

This paper cites Trajectory optimization for nonlinear multi-agent systems using decentralized learning model predictive control.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL Trajectory optimization for nonlinear multi-agent systems using decentralized learning model predictive control

Reference 92

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:35:49.990352Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-16T11:35:49.560635Z digest=sha256:4367efaa2372066b29e589302a8031e355ddd8b3c5f62089f9d3bb52fbea5f9f

Observation 26b1b56b-96f7-4307-bc1d-b75711f7a03f · outbound

This paper cites In other words, for a given z0, the value at the kth timestep is only a function of zk and xk instead of the z0 and the entire trajectory up to the kth timestep.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL In other words, for a given z0, the value at the kth timestep is only a function of zk and xk instead of the z0 and the entire trajectory up to the kth timestep

Reference 93

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:35:49.972993Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-16T11:35:49.565966Z digest=sha256:cb269140f1b6013d31dcf1beb8ada9e561f96e5a5a7f0648d0bfdeacbe44ebb7

Observation f0cdaa1e-0e66-4000-aae9-24dc084b845f · outbound

This paper cites an unresolved cited work.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL Unresolved cited work

Reference 94

Resolution
unresolved
raw_fallback, observed 2026-08-16T11:35:49.955745Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-16T11:35:49.570731Z digest=sha256:cfc608eb92860b220e901002628724f11499f5c8aea66a53166e3be26a3abd12

Observation 6cffaff4-cbbf-4eac-8f7f-8c27fe1a09b2 · outbound

This paper cites an unresolved cited work.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL Unresolved cited work

Reference 95

Resolution
unresolved
raw_fallback, observed 2026-08-16T11:35:49.938382Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-16T11:35:49.575826Z digest=sha256:10fe16ef4064130aad95acbeea2c580b69c3457467fe6155e1b4410b58b33f12

Observation 53d0bdd6-665c-4045-a16d-ea6f6b03ab4d · outbound

This paper cites an unresolved cited work.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL Unresolved cited work

Reference 96

Resolution
unresolved
raw_fallback, observed 2026-08-16T11:35:49.922510Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-16T11:35:49.580801Z digest=sha256:4e78b221144885e0f8ad232a5ce854b361f89cd236a304d3dfe67a6a3a359cd1

Observation be1d4854-de22-48a7-876d-7e1571c4931d · outbound

This paper cites APPENDIX D ALGORITHM PSEUDOCODE We describe the centralized training process of Def-MARL in Algorithm 1 and the distributed execution process in Algorithm 2.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL APPENDIX D ALGORITHM PSEUDOCODE We describe the centralized training process of Def-MARL in Algorithm 1 and the distributed execution process in Algorithm 2

Reference 97

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:35:49.907043Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-16T11:35:49.585636Z digest=sha256:97a56faba7848df64ba8b451e0d5e58d24157003d9b8a888b9cbbe8668700186

Observation 99cdc52b-244d-4645-aa18-6dae73dd4be0 · outbound

This paper cites E⊆{ (i,j )|i∈V a,j ∈V} is the set of edges, denoting the information flow from a sender node j to a receiver agent i.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL E⊆{ (i,j )|i∈V a,j ∈V} is the set of edges, denoting the information flow from a sender node j to a receiver agent i

Reference 98

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:35:49.889638Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-16T11:35:49.590785Z digest=sha256:c5a352f7a3a0810c52eeaafe17b2d865e3eed20b4474c06e1a3ae72040b65218

Observation 2884d2e1-7170-4d80-beb3-d044cd74f8c9 · outbound

This paper cites an unresolved cited work.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL Unresolved cited work

Reference 99

Resolution
unresolved
raw_fallback, observed 2026-08-16T11:35:49.870781Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-16T11:35:49.596258Z digest=sha256:568c78d05b66f41005f794d43f5d5811909603ba45577b96a6fb6c9b5faea741

Pith citing papers

Observation e8d1989e-45f5-44b0-bf8c-d9b029676de5 · inbound

Distributed Safety-Critical Control of Multi-Agent Systems with Time-Varying Communication Topologies cites this paper.

Distributed Safety-Critical Control of Multi-Agent Systems with Time-Varying Communication Topologies Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL

Reference 1

Resolution
verified exact
arxiv_id, observed 2026-05-13T22:58:23.933705Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-05-13T22:56:33.073457Z digest=sha256:c38cb89d7286cf44e940fd4425ba7a2a998c77fc185c507ac93bfe3a4e0295d4