Pith. sign in

Paper Citation Record · LEDGER

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL

As of 17 August 2026, this Paper Citation Record lists 99 of 99 outbound references and 1 inbound Pith citation observation for arXiv:2504.15425.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2504.15425 v1

Coverage vector

measured 99 of 99 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-16T11:35:49.596258Z

measured 100 of 100 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-17T06:30:58.91139+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-05-13T22:56:33.073457Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-13T22:58:23.930563Z

Reference resolution

99 of 99 outbound references displayed

  • verified exact0
  • verified fuzzy56
  • unresolved43
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 021ed894-e2a8-4fad-af50-233bc9b42bc9 · outbound

This paper cites Constrained policy optimization.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL Constrained policy optimization

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-16T11:35:49.100505Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:35:49.100505Z digest=sha256:d252c26344163f29c6651c5dcba97b745f365e5e74e181642a2762e31ba9e1e1

Observation 3f762109-487a-40e7-a875-1da5a0add65d · outbound

This paper cites Learning transferable cooperative be- havior in multi-agent team.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL Learning transferable cooperative be- havior in multi-agent team

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-16T11:35:49.105943Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:35:49.105943Z digest=sha256:ed027b860301e5e5b13afca74ab3bfdc168d232a204cd5ff7d2cb244d4b6222b

Observation f011d933-34b0-4ec8-956f-40751c62e234 · outbound

This paper cites Constrained Markov decision processes.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL Constrained Markov decision processes

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-16T11:35:49.111335Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:35:49.111335Z digest=sha256:e1f1bb8514499bec59b9bec81c1d734869821009b32474dd10a70cb9c9d854d2

Observation bddfa60f-aab9-4940-a474-151512c7062a · outbound

This paper cites Casadi: a software framework for nonlinear optimization and optimal con- trol.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL Casadi: a software framework for nonlinear optimization and optimal con- trol

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-16T11:35:49.116234Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:35:49.116234Z digest=sha256:3605eb0fc22e5edb2b384873818941b70e05c40fcce1d20092ceff4fc2eb70c4

Observation ba6db136-ebf3-4e72-8c1b-16766ba2e162 · outbound

This paper cites Hamilton-jacobi reachability: A brief overview and recent advances.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL Hamilton-jacobi reachability: A brief overview and recent advances

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-16T11:35:49.121126Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:35:49.121126Z digest=sha256:fc37d29c60475e6d29e9a1b3b18d570b1c17f664578b16c838ad04abfb2102c7

Observation 66e5c626-3936-4ea5-98fb-f527b5901967 · outbound

This paper cites Dynamic programming and optimal control: Volume I, volume 4.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL Dynamic programming and optimal control: Volume I, volume 4

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-16T11:35:49.125942Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:35:49.125942Z digest=sha256:5667976b69d335126857dff161f7b6c706fb7222f4050860ce5922355e5decd5

Observation 647f9e09-641c-4aa1-b5cb-6078b0a34493 · outbound

This paper cites Synthesis of minimum-cost shields for multi-agent systems.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL Synthesis of minimum-cost shields for multi-agent systems

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-16T11:35:49.131262Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:35:49.131262Z digest=sha256:9100e606d093db257cf5afec31f730a2236186098c1ed9ad23d5addaa04c3d22

Observation 01d8c8e2-5ed1-4d53-b4da-b702c9846195 · outbound

This paper cites An actor-critic algorithm for constrained markov decision processes.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL An actor-critic algorithm for constrained markov decision processes

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-16T11:35:49.136145Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:35:49.136145Z digest=sha256:e635faa268484a275ab4a81d9c05c99f9332d768c0c3624e5e127456982a301a

Observation 387cca71-61aa-4619-9945-16a5dd173f73 · outbound

This paper cites Stochastic Approximation: A Dynamical Systems Viewpoint, volume 48.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL Stochastic Approximation: A Dynamical Systems Viewpoint, volume 48

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-16T11:35:49.141464Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:35:49.141464Z digest=sha256:0057551c45291a9893430b5496c0083c0fd094ec0593a41f76e9c472ae7abe19

Observation f2eede24-1a56-4d49-8870-c9075d9289b3 · outbound

This paper cites Convex optimization.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL Convex optimization

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-16T11:35:49.146263Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:35:49.146263Z digest=sha256:1b819479d1be03b4bd6834a4be33356ceee9afee32b76e2969e3edd02d1b6c00

Observation b3dae811-fdde-4419-b16c-879328c11f8f · outbound

This paper cites Safe Multi-Agent Reinforcement Learning through Decentralized Multiple Control Barrier Functions.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL Safe Multi-Agent Reinforcement Learning through Decentralized Multiple Control Barrier Functions

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-16T11:35:49.151496Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:35:49.151496Z digest=sha256:34914bd4a058f26cabe6ad4cff7ebd1dd6cf403add324e081ee6156de2ed4721

Observation 47be9f9e-05d5-44dd-8c33-ac89b9305b93 · outbound

This paper cites A new hybrid quadratic/bisection algorithm for finding the zero of a nonlinear function without using derivatives.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL A new hybrid quadratic/bisection algorithm for finding the zero of a nonlinear function without using derivatives

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-16T11:35:49.156876Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:35:49.156876Z digest=sha256:67011cd3ba72930a960909c0332dbd5ef4dc65bff0f7fb7696ad026d387324eb

Observation 1182b6e9-eee5-4d80-9d2c-bf4c05b5cc3c · outbound

This paper cites Socially aware motion planning with deep re- inforcement learning.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL Socially aware motion planning with deep re- inforcement learning

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-16T11:35:49.161771Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:35:49.161771Z digest=sha256:b326335c504771c6a2151f547be1446870469e58e9a592f25756e47273cfc72a

Observation 4924cc63-5136-465e-8dee-b868114ab8ec · outbound

This paper cites Decentralized non-communicating multiagent col- lision avoidance with deep reinforcement learning.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL Decentralized non-communicating multiagent col- lision avoidance with deep reinforcement learning

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-16T11:35:49.166199Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:35:49.166199Z digest=sha256:c7561c252c62fb6bc9d277a55eac59cca8bc7c1beed42c2471499accb44f7130

Observation 2986275c-4cda-40fe-b0ac-f6f90ac66d74 · outbound

This paper cites On the duality gap of constrained cooperative multi-agent reinforcement learning.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL On the duality gap of constrained cooperative multi-agent reinforcement learning

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-16T11:35:49.170705Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:35:49.170705Z digest=sha256:3a679b69fab7f4cc0d5b4fb219ca66cf198d86ef887d061523ab3e2b631a1fd6

Observation dfd586fa-ea56-4f86-b855-92705b94571f · outbound

This paper cites Computational aspects of distributed optimization in model predictive control.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL Computational aspects of distributed optimization in model predictive control

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-16T11:35:49.175559Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:35:49.175559Z digest=sha256:47729d23e4d73fe0dfe9f72d988737d080f5413a2d1a8358fa079d418b1f131c

Observation e1c51361-7511-40ce-b559-ece92db3f9bd · outbound

This paper cites De- tecting, localizing, and tracking an unknown number of moving targets using a team of mobile robots.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL De- tecting, localizing, and tracking an unknown number of moving targets using a team of mobile robots

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-16T11:35:49.180383Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:35:49.180383Z digest=sha256:c2716b5e744a7bf23752390fcb66ecc5a2a2c87596041705d89b6c3c682064f6

Observation 357a9e3c-ec56-4522-860a-ca97b7967865 · outbound

This paper cites Provably efficient gener- alized lagrangian policy optimization for safe multi-agent reinforcement learning.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL Provably efficient gener- alized lagrangian policy optimization for safe multi-agent reinforcement learning

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-16T11:35:49.185205Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:35:49.185205Z digest=sha256:85cd3e034fbe88e82454127ab68211a113a9f2d25df1fd1b64c13fc4a5219703

Observation 81d691fe-8670-41e6-b37e-f155dddd9993 · outbound

This paper cites Safe Multi-Agent Reinforcement Learning via Shielding.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL Safe Multi-Agent Reinforcement Learning via Shielding

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-16T11:35:49.190136Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:35:49.190136Z digest=sha256:dfc29ead2c7974a3e13ca1f8b9c8466c56ec4b5d1fed443de3881a474e295cb1

Observation e0024367-a868-409e-8988-6a3c1dbe99a2 · outbound

This paper cites Safe multi- agent reinforcement learning via shielding.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL Safe multi- agent reinforcement learning via shielding

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-16T11:35:49.194900Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:35:49.194900Z digest=sha256:edd9cad656325432509b9faa49516a477acbe9d451b1243d00b717e3bdd02524

Observation 5101e1bb-d11f-46c4-b075-44e424cad64b · outbound

This paper cites Mo- tion planning among dynamic, decision-making agents with deep reinforcement learning.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL Mo- tion planning among dynamic, decision-making agents with deep reinforcement learning

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-16T11:35:49.199462Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:35:49.199462Z digest=sha256:91e66e2993144a6c1d0bba8d4e88210e41a98fc498a016f558796b74cddddbe6

Observation a2625dc9-4098-48a1-a000-2456fb48fdf1 · outbound

This paper cites A distributed model predictive control strategy for constrained multi- agent systems: The uncertain target capturing scenario.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL A distributed model predictive control strategy for constrained multi- agent systems: The uncertain target capturing scenario

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:35:50.960217Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-16T11:35:49.203864Z digest=sha256:5c4e6db9909a36c26e5366b122beaee6df81587ddb3e75b40a65c7b624a26eac

Observation a37f3fec-48f4-496b-9e91-42bd4d4dca69 · outbound

This paper cites Counterfactual multi-agent policy gradients.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL Counterfactual multi-agent policy gradients

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:35:50.944198Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-16T11:35:49.208585Z digest=sha256:638e71d39be280d78619d0c7573be386585d90caeb3f9c5e816710ee9b36e663

Observation a36e666d-7a33-4222-aba9-4073d813675b · outbound

This paper cites Iterative reachability estimation for safe reinforcement learning.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL Iterative reachability estimation for safe reinforcement learning

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-16T11:35:49.213377Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:35:49.213377Z digest=sha256:d9097eaa0800e63826f1ccf4a0cfaf1769db1271707117c6ed8433cad58c1691

Observation 3ca00963-31c1-42f3-99d1-b440454e11c5 · outbound

This paper cites Learning safe control for multi- robot systems: Methods, verification, and open chal- lenges.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL Learning safe control for multi- robot systems: Methods, verification, and open chal- lenges

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:35:50.916330Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-16T11:35:49.218177Z digest=sha256:d2f0c045061d45b79e7fc83f4d46798e4fd2e858eed7b1618b2d7b09f9ba8932

Observation 2e19ab2a-784e-41e9-9806-908398d7a7fc · outbound

This paper cites A reinforce- ment learning framework for vehicular network routing under peak and average constraints.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL A reinforce- ment learning framework for vehicular network routing under peak and average constraints

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:35:50.899242Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-16T11:35:49.222823Z digest=sha256:6c6fbac0930a9fdf39667da0261724cfd04636b1854f6cfd9a3ae0a80ee1a9f4

Observation 7e679174-932b-4bb9-9733-7274dc2c81d5 · outbound

This paper cites Crazyflie 2.0 quadrotor as a platform for research and education in robotics and control engineering.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL Crazyflie 2.0 quadrotor as a platform for research and education in robotics and control engineering

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:35:50.882587Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-16T11:35:49.228778Z digest=sha256:3f7f3c6f3f478768e4cd404341b815ae406ebcac7dbe85af8bb404e0b030ef63

Observation 6f6e2a00-a966-4689-a670-3e1230b068cd · outbound

This paper cites Snopt: An sqp algorithm for large-scale constrained optimization.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL Snopt: An sqp algorithm for large-scale constrained optimization

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-16T11:35:49.234123Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:35:49.234123Z digest=sha256:95ca8a6e9716484691dd1d88ee3b7b1e901140a3edbd1dce4fcda1466c981531

Observation c7c38f09-8592-41a2-b2e7-97f7db9891a4 · outbound

This paper cites Nonlinear model predictive control: theory and algorithms.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL Nonlinear model predictive control: theory and algorithms

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:35:50.853203Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-16T11:35:49.239025Z digest=sha256:c93788afa87b358327b9a043c5eaef01df963f48f1b09017b7cf28f99b3ff748

Observation cd1c5b7b-32eb-4073-a64f-0fea3ec9c907 · outbound

This paper cites Multi-Agent Constrained Policy Optimisation.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL Multi-Agent Constrained Policy Optimisation

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-16T11:35:49.243705Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:35:49.243705Z digest=sha256:e183acbf46b8e4d1438bc69ec4b520c1b17f486921a3df547a141045ebf13abf

Observation 75bcca73-1d38-4d39-b5ae-6474eddf8218 · outbound

This paper cites A Review of Safe Reinforcement Learning: Methods, Theory and Applications.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL A Review of Safe Reinforcement Learning: Methods, Theory and Applications

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-16T11:35:49.248824Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:35:49.248824Z digest=sha256:d4e9bacff836c54bcb6f5085a65bd5e3b02b8cb4ff9a2ee7986f2b2ae9efc19c

Observation df65dcad-5e97-4d33-84b7-8389674db06d · outbound

This paper cites Safe multi-agent reinforcement learning for multi-robot control.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL Safe multi-agent reinforcement learning for multi-robot control

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:35:50.834594Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-16T11:35:49.253790Z digest=sha256:8a36577623c721280f55ac365062115adb1d8cfd164f6b485e5389bd120298cc

Observation 7fe7dc21-ca19-41c5-a1c5-f1443941b0a2 · outbound

This paper cites Coordinated reinforcement learning.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL Coordinated reinforcement learning

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:35:50.816795Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-16T11:35:49.258380Z digest=sha256:1a035fc139a59ca58d7e5d45a977067fc1ec58b9aec86f640816982eaaf0c3a1

Observation bce72a09-d609-4b10-9671-4cce7b9d10ed · outbound

This paper cites Deep recurrent q-learning for partially observable mdps.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL Deep recurrent q-learning for partially observable mdps

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:35:50.800028Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-16T11:35:49.263244Z digest=sha256:83b249dfdcb3cd2156ae591c14ac83d69252b2ae8904ca8d7cf83dde97252666

Observation efc781f4-8330-4f9e-8ab0-b0fa59a789d6 · outbound

This paper cites Autocost: Evolving intrinsic cost for zero-violation reinforcement learning.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL Autocost: Evolving intrinsic cost for zero-violation reinforcement learning

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:35:50.782922Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-16T11:35:49.269034Z digest=sha256:3667270ef617cb2fd747ef67f68297d26f142463a828898c05c04f1148fedf72

Observation 60b7bea2-7cf5-4183-ab78-6e180f5141a5 · outbound

This paper cites Safedreamer: Safe reinforcement learning with world models.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL Safedreamer: Safe reinforcement learning with world models

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:35:50.765390Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-16T11:35:49.273968Z digest=sha256:eb124122a97f9dfc3aac0e2e24eab1c16d3986cfe42efd99fd9ab76aa86ac82d

Observation bcef9d51-07e5-449c-8798-89fdffd4bdb4 · outbound

This paper cites Distributed optimization in multi-agent robotics for industry 4.0 warehouses.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL Distributed optimization in multi-agent robotics for industry 4.0 warehouses

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:35:50.747003Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-16T11:35:49.280957Z digest=sha256:f6e11f46c3d3b92793807789ddd527a5ad2238286dd14ef01635c54e3b714138

Observation ea097793-3e3b-4905-be52-a3a255c5c929 · outbound

This paper cites Cmix: Deep multi- agent reinforcement learning with peak and average con- straints.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL Cmix: Deep multi- agent reinforcement learning with peak and average con- straints

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:35:50.728160Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-16T11:35:49.286188Z digest=sha256:76d4f375e9af1299d623f131f2152ff46b2c027d14631fe33ec1b1fecf2091ec

Observation 1406ba9c-9992-4145-9e9c-3710efa7d9fc · outbound

This paper cites Towards optimally decentralized multi-robot collision avoidance via deep reinforcement learning.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL Towards optimally decentralized multi-robot collision avoidance via deep reinforcement learning

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:35:50.705889Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-16T11:35:49.290715Z digest=sha256:c05622f137d7295a49713122bd409370ba787d64504e31a865fa355805587edc

Observation 08f7fcea-19d9-480d-a403-5c4ed1d81e61 · outbound

This paper cites Multi-agent actor- critic for mixed cooperative-competitive environments.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL Multi-agent actor- critic for mixed cooperative-competitive environments

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:35:50.687627Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-16T11:35:49.295450Z digest=sha256:59a2417e808f0e833270411be1359772bb645f5c973f1b7431bd6052fb893dae

Observation 1ec2c0d6-124d-4e49-8d1c-7e4796f40a9c · outbound

This paper cites Decentralized policy gradient descent ascent for safe multi-agent reinforcement learning.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL Decentralized policy gradient descent ascent for safe multi-agent reinforcement learning

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:35:50.670639Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-16T11:35:49.300993Z digest=sha256:0ea131f9a5ada4a7f59a32be26c4b640d58c27505f96a51ce938ff5bb6dae91a

Observation 37f2e7e6-78ba-46cb-8ac7-34d294efe1ca · outbound

This paper cites Trajectory generation for multiagent point-to-point transitions via distributed model predictive control.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL Trajectory generation for multiagent point-to-point transitions via distributed model predictive control

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:35:50.653222Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-16T11:35:49.306462Z digest=sha256:73e5b46156de129d2a93c14afdea3c4d72867cef2708e1e2e7988d016f094712

Observation 94b03caa-0cb1-4e78-8c51-3b1b3280dc16 · outbound

This paper cites Online trajectory generation with distributed model predictive control for multi-robot motion planning.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL Online trajectory generation with distributed model predictive control for multi-robot motion planning

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:35:50.634565Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-16T11:35:49.311335Z digest=sha256:f89c7455c0e688a30a03238512e900664db53c91094ceeb779b304f4c68b5f7b

Observation 3d62944b-59f5-43d1-a914-1388f859acdf · outbound

This paper cites On reachability and minimum cost optimal control.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL On reachability and minimum cost optimal control

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-16T11:35:49.316420Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:35:49.316420Z digest=sha256:a5be2abcd4fb366f7e0157298e32c5d0976511fc5b27a8c909e04f6fd05af856

Observation a67092fb-f45b-4168-8ce2-1c802594f8d9 · outbound

This paper cites Lifelong Multi-Agent Path Finding for Online Pickup and Delivery Tasks.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL Lifelong Multi-Agent Path Finding for Online Pickup and Delivery Tasks

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-16T11:35:49.321247Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:35:49.321247Z digest=sha256:6012279fc0cece438fc802e3b1f5091ed9fd850172e1f09abe7baee6de6abe0b

Observation 93311ed5-1199-4c60-9bf5-029f404aa922 · outbound

This paper cites Hamilton–jacobi formulation for reach–avoid differential games.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL Hamilton–jacobi formulation for reach–avoid differential games

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:35:50.605556Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-16T11:35:49.326900Z digest=sha256:67540d45a119764e8870a533d45a853a6099a00f6588d52d572b11df39e841ca

Observation bfd02cfa-a22b-4953-be02-47eba0a8793a · outbound

This paper cites Safe value functions.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL Safe value functions

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:35:50.587653Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-16T11:35:49.331833Z digest=sha256:af7b2ad16952a893126d655bc7652d733ea4547ac1bff0af00907d13fc9b3a39

Observation bc854208-b398-4db6-beec-579021b599d8 · outbound

This paper cites Shield decentralization for safe multi-agent reinforce- ment learning.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL Shield decentralization for safe multi-agent reinforce- ment learning

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:35:50.570150Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-16T11:35:49.337215Z digest=sha256:5cdbd2a693772fa662cbf16ccf183d342023dfd60c5a33ceab28bbe639132ed4

Observation 140a3280-0ba2-4594-8868-3c5926e56dfd · outbound

This paper cites A time-dependent hamilton-jacobi formulation of reachable sets for continuous dynamic games.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL A time-dependent hamilton-jacobi formulation of reachable sets for continuous dynamic games

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-16T11:35:49.342067Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:35:49.342067Z digest=sha256:24bdb714dedce8a33de6a041c1de0e0762c7b909a5064dc0fb591f33a1378e5e

Observation e0fec4a7-ff8b-4d82-a8bc-93d754141da9 · outbound

This paper cites Distributed model predictive safety certification for learning-based control.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL Distributed model predictive safety certification for learning-based control

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:35:50.540428Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-16T11:35:49.346763Z digest=sha256:de80e8703f219bf1435cd851f27799a43a9c16aeccbdfb4bcac0bf428c09826b

Observation 2f02698a-e874-44db-974c-e81feffb43b8 · outbound

This paper cites Scalable multi-agent reinforcement learning through intelligent information aggregation.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL Scalable multi-agent reinforcement learning through intelligent information aggregation

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:35:50.523861Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-16T11:35:49.351927Z digest=sha256:f18fdc24db59b73bc502e90a60866b524d70d3ed5ee37d0d514d7295623ba999

Observation fd4865da-e1d0-4b21-a8d3-347e590770d3 · outbound

This paper cites Distributed optimization for control.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL Distributed optimization for control

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:35:50.506329Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-16T11:35:49.356555Z digest=sha256:37618ebcbb5886d21bf1ae081ad1b4738b2d60aaecfbcd9c4858d26e9c16c583

Observation a9691ad1-2670-49cf-b45e-2c22cdba0e9c · outbound

This paper cites Numerical opti- mization.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL Numerical opti- mization

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-16T11:35:49.361981Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:35:49.361981Z digest=sha256:e527ae701a2b43432259db6d482b5ffa0e1432a9f342039d961404c2f7824230

Observation 13192686-87fe-4a3b-9a5c-6f1a74d3a50f · outbound

This paper cites Facmac: Factored multi- agent centralised policy gradients.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL Facmac: Factored multi- agent centralised policy gradients

Reference 54

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:35:50.479369Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-16T11:35:49.366973Z digest=sha256:29c6a9f15f6e226caea4ca7e9f94233ad756599b0d54d76c9b5fd5ace854b095

Observation 21340052-cdf4-4616-b3b5-61c0855e9881 · outbound

This paper cites Decentralized Safe Multi-agent Stochastic Optimal Control using Deep FBSDEs and ADMM.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL Decentralized Safe Multi-agent Stochastic Optimal Control using Deep FBSDEs and ADMM

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-16T11:35:49.372051Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:35:49.372051Z digest=sha256:97d5c817dc1777e159ee19bb662c0f17f29477d9a360123acbae1be0cad875b6

Observation aabf5350-5c2b-4d98-8168-1bd78439b3fb · outbound

This paper cites Learning safe multi-agent control with decentralized neural barrier certificates.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL Learning safe multi-agent control with decentralized neural barrier certificates

Reference 56

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:35:50.462917Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-16T11:35:49.377313Z digest=sha256:7f34964b15fab4f29a6399acde3c1f0cea6564aa2278a1219ece1ff91b76eeac

Observation f2d235d6-6442-4c49-b465-007add8b0b91 · outbound

This paper cites Weighted qmix: Expanding monotonic value function factorisation for deep multi-agent reinforcement learning.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL Weighted qmix: Expanding monotonic value function factorisation for deep multi-agent reinforcement learning

Reference 57

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:35:50.446467Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-16T11:35:49.382001Z digest=sha256:d8116f53d6136e172ec4348f7127578c62868ca8f017163c86692b64dbce1389

Observation 9146e6a4-38c0-47b9-b853-b50494f6c11c · outbound

This paper cites Monotonic value function factorisation for deep multi-agent reinforcement learning.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL Monotonic value function factorisation for deep multi-agent reinforcement learning

Reference 58

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:35:50.430474Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-16T11:35:49.386675Z digest=sha256:2b929671091e0064b545dfbdbb39eb60ceb9b315c73bec22ed004a3a15cb7f5c

Observation bd3fb1fe-ecc5-4994-8921-57747eb556e1 · outbound

This paper cites A stochastic approx- imation method.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL A stochastic approx- imation method

Reference 59

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:35:50.414429Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-16T11:35:49.391565Z digest=sha256:d25df7e32acc8aa961fe33e2c1137673939bee019cc259a8a37cc5c9f6192a5f

Observation 3c9c3079-afe6-4d49-845b-79ebf0a70db0 · outbound

This paper cites Con- strained markov decision processes via backward value functions.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL Con- strained markov decision processes via backward value functions

Reference 60

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:35:50.398173Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-16T11:35:49.396466Z digest=sha256:3769f9db9e75ae454d1479697560e76bc7cf6ae9765c62fc6d0e80b9170f047f

Observation d6075347-e324-4dba-8847-a101ca763e47 · outbound

This paper cites Trust region policy optimiza- tion.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL Trust region policy optimiza- tion

Reference 61

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:35:50.381182Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-16T11:35:49.401269Z digest=sha256:0beb9e08e11ea680f9e92fbf536c4b3872d66b1bfda6904378e464abb0e230ad

Observation 3b10c4ad-aecd-42e6-93e6-5b509e001dc4 · outbound

This paper cites High-Dimensional Continuous Control Using Generalized Advantage Estimation.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL High-Dimensional Continuous Control Using Generalized Advantage Estimation

Reference 62

Resolution
unresolved
no resolver link, observed 2026-08-16T11:35:49.405908Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:35:49.405908Z digest=sha256:f6365b7fe3254b50b05f88fd23f2204420fde090f77eddb09b2d084d6f438eda

Observation fee63235-7e9a-4da2-a553-945bc6e6ac5b · outbound

This paper cites Proximal Policy Optimization Algorithms.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL Proximal Policy Optimization Algorithms

Reference 63

Resolution
unresolved
no resolver link, observed 2026-08-16T11:35:49.411092Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:35:49.411092Z digest=sha256:61509e8010eb60e4d1d38cad97b9bbe3b8422d52af646490a83eea6f5909854b

Observation 3adb61f4-599e-4493-8fef-ffab520f8d3c · outbound

This paper cites Multi-agent motion planning for dense and dynamic environments via deep reinforcement learning.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL Multi-agent motion planning for dense and dynamic environments via deep reinforcement learning

Reference 64

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:35:50.363450Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-16T11:35:49.415932Z digest=sha256:83eac5977181c9ff6562240de154f26f860ec66939ad832495cb2e9f2783a828

Observation 41fad7e6-d630-4b0c-ba21-c3394c99f9dd · outbound

This paper cites Masked Label Prediction: Unified Message Passing Model for Semi-Supervised Classification.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL Masked Label Prediction: Unified Message Passing Model for Semi-Supervised Classification

Reference 65

Resolution
unresolved
no resolver link, observed 2026-08-16T11:35:49.420858Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:35:49.420858Z digest=sha256:c7a014dc86d5de1bcf4bebafd0c5ba349032f7c38ccbcbe378deb860d46bbee4

Observation 1309f659-a663-49a7-8ca9-119b5083d7bb · outbound

This paper cites Solving stabilize-avoid optimal control via epigraph form and deep reinforce- ment learning.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL Solving stabilize-avoid optimal control via epigraph form and deep reinforce- ment learning

Reference 66

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:35:50.347386Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-16T11:35:49.426325Z digest=sha256:380c9817f8ec8cd8eecacd3e477fbb507c18890c734cd6fd60c2ad871105a2dc

Observation 29aebdd6-f22a-4bed-ab12-78370cd16582 · outbound

This paper cites Solving minimum-cost reach avoid using reinforcement learning.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL Solving minimum-cost reach avoid using reinforcement learning

Reference 67

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:35:50.330772Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-16T11:35:49.431478Z digest=sha256:f6c106d6dfdca9ee599f65cdc8b795cb09ed249314aa3defe0d80a6e55c81408

Observation bf9028ad-636b-40e7-baee-5aa78a81ff94 · outbound

This paper cites Predictive control of aerial swarms in cluttered environ- ments.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL Predictive control of aerial swarms in cluttered environ- ments

Reference 68

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:35:50.311938Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-16T11:35:49.436382Z digest=sha256:403e90ad6c2dfa14f86d79b9299c98cb5ef1afcc2fc39249227aae4929101084

Observation 1926dfd9-129d-44d3-9ebb-4915cc2c9790 · outbound

This paper cites Value-Decomposition Networks For Cooperative Multi-Agent Learning.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL Value-Decomposition Networks For Cooperative Multi-Agent Learning

Reference 69

Resolution
unresolved
no resolver link, observed 2026-08-16T11:35:49.441863Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:35:49.441863Z digest=sha256:5ad4efcd0692cce1484f10c2911e7515f972e7f8492c7fd41e16d38e69a5a7de

Observation 54313299-44bc-4c9e-ba05-01afdf70c610 · outbound

This paper cites Mankowitz, and Shie Mannor.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL Mankowitz, and Shie Mannor

Reference 70

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:35:50.293658Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-16T11:35:49.447051Z digest=sha256:dca77d0cbbad2df05962a1db27b617244474f8e10eb6f07b733b6a5bac867b67

Observation 1522d750-8cc9-46c0-92a8-36bb1816632f · outbound

This paper cites A game theoretic approach to controller design for hybrid systems.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL A game theoretic approach to controller design for hybrid systems

Reference 71

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:35:50.276454Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-16T11:35:49.452061Z digest=sha256:be5ed07cccc1d94dd8cf1fa5dac86d26cf056b26d52e2222f4166aebc842606e

Observation fcfdc6fa-dcc9-44db-82af-7ef417761499 · outbound

This paper cites Decentralized multi-agent planning using model predictive control and time-aware safe corridors.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL Decentralized multi-agent planning using model predictive control and time-aware safe corridors

Reference 72

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:35:50.260275Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-16T11:35:49.456929Z digest=sha256:422fe75b1760cab1d2590d28b4d823394e23e864d922d11704aceb7bd416a505

Observation f2d69c81-cba1-4274-9b0a-19aa5399a1d1 · outbound

This paper cites Initial guess generation for aircraft landing trajec- tory optimization.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL Initial guess generation for aircraft landing trajec- tory optimization

Reference 73

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:35:50.241930Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-16T11:35:49.461859Z digest=sha256:a4a46d6f78ae7810c4bfa025ddbffb595caefc953ca2a2258d8267a2acf4ab10

Observation 85c5ed1a-92ab-48d1-ad9b-02b126e0355f · outbound

This paper cites QPLEX: Duplex Dueling Multi-Agent Q-Learning.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL QPLEX: Duplex Dueling Multi-Agent Q-Learning

Reference 74

Resolution
unresolved
no resolver link, observed 2026-08-16T11:35:49.466812Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:35:49.466812Z digest=sha256:1d5b026e29cf5e8eec708da50f59541661f51d506a635a442bb7564b23b5b941

Observation 587b983f-2f82-460d-b4f8-21478fb37347 · outbound

This paper cites A synthesis approach of distributed model predictive control for homogeneous multi-agent system with collision avoidance.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL A synthesis approach of distributed model predictive control for homogeneous multi-agent system with collision avoidance

Reference 75

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:35:50.225310Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-16T11:35:49.472565Z digest=sha256:a5df1d07c8a1e2f8f468bfd872880f2aa91944249d8cdde2ae61b3180cef267f

Observation 04916c6e-b7d6-41bc-be07-e2a7093bf891 · outbound

This paper cites Multi-agent deep reinforcement learning for urban traffic light control in vehicular networks.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL Multi-agent deep reinforcement learning for urban traffic light control in vehicular networks

Reference 76

Resolution
unresolved
no resolver link, observed 2026-08-16T11:35:49.478723Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:35:49.478723Z digest=sha256:11756db00305e7bd42bc7e235caeb7d4f8a05bfd8ebeb4d70952526008782643

Observation abafd294-d515-4529-a413-1c0fd57fab81 · outbound

This paper cites Model-based Dynamic Shielding for Safe and Efficient Multi-Agent Reinforcement Learning.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL Model-based Dynamic Shielding for Safe and Efficient Multi-Agent Reinforcement Learning

Reference 77

Resolution
unresolved
no resolver link, observed 2026-08-16T11:35:49.483573Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:35:49.483573Z digest=sha256:7240b7338aef9d4ebf7de0f436a35bbcd95c115454855e4615b740472e7cb1b6

Observation 1f6914fa-a1f2-4d92-8206-66da2e5e14e4 · outbound

This paper cites Crpo: A new approach for safe reinforcement learning with convergence guarantee.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL Crpo: A new approach for safe reinforcement learning with convergence guarantee

Reference 78

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:35:50.196643Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-16T11:35:49.488687Z digest=sha256:7399a809158313eec54cd24e0226805567428abd83a693f1d4056ec03a2c1390

Observation 71321466-c55e-4511-a348-002b28d81941 · outbound

This paper cites Qatten: A General Framework for Cooperative Multiagent Reinforcement Learning.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL Qatten: A General Framework for Cooperative Multiagent Reinforcement Learning

Reference 79

Resolution
unresolved
no resolver link, observed 2026-08-16T11:35:49.493384Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:35:49.493384Z digest=sha256:1f87c26864d034159dd29cc11b178e439e79280b720d74221e4c58bdb09ef555

Observation 69237a14-51b7-4de9-a903-c724df5cc2e1 · outbound

This paper cites The surprising effectiveness of ppo in cooperative multi-agent games.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL The surprising effectiveness of ppo in cooperative multi-agent games

Reference 80

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:35:50.178088Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-16T11:35:49.499001Z digest=sha256:f5359682baef5ae2c9eeac09076fec31eaa183fa8b6706cb47f934672afe22b6

Observation 9095540a-b7ea-4b7b-9411-59b023a5f2e6 · outbound

This paper cites Reachability constrained reinforcement learning.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL Reachability constrained reinforcement learning

Reference 81

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:35:50.161635Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-16T11:35:49.504513Z digest=sha256:5d9cdba914a6f86de07097e2f74b9b0c6a8d22d803c1417c4be44924ae79d8cb

Observation b2c0f567-268e-4b73-9e38-2d1f76ccf2b8 · outbound

This paper cites Safe reinforcement learning using robust mpc.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL Safe reinforcement learning using robust mpc

Reference 82

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:35:50.143568Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-16T11:35:49.509472Z digest=sha256:0aa3859ac5674ca3dead17c417da031aeaac265b7d93a2c9758004783a0dc2b8

Observation 73060d35-0d38-480d-bfe0-fee0857b7def · outbound

This paper cites Fully decentralized multi-agent re- inforcement learning with networked agents.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL Fully decentralized multi-agent re- inforcement learning with networked agents

Reference 83

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:35:50.124360Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-16T11:35:49.514827Z digest=sha256:5fe939722a844ce6d18758ccbb74ff622bdd0c1dd6c0e7453044085b44bda191

Observation ffe68379-54e4-4d0b-a739-3bd8773eda9e · outbound

This paper cites Multi- agent reinforcement learning: A selective overview of theories and algorithms.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL Multi- agent reinforcement learning: A selective overview of theories and algorithms

Reference 84

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:35:50.105282Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-16T11:35:49.520678Z digest=sha256:d381f6fb9bcbfdd5f0d996c4982c37b947ac42448dafa31a8180bf390c941d37

Observation 6eb47d82-ce9b-450e-a34e-ae2cb5161338 · outbound

This paper cites Neu- ral graph control barrier functions guided distributed collision-avoidance multi-agent control.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL Neu- ral graph control barrier functions guided distributed collision-avoidance multi-agent control

Reference 85

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:35:50.088952Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-16T11:35:49.525301Z digest=sha256:4e576e2059e7c62e981ac0bfc17f3783c394af5840efed3b0aaf78823b3e280b

Observation cc55ae64-209c-4442-a094-ae8e0610b2d1 · outbound

This paper cites Discrete GCBF proximal policy optimization for multi-agent safe optimal control.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL Discrete GCBF proximal policy optimization for multi-agent safe optimal control

Reference 86

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:35:50.072861Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-16T11:35:49.530170Z digest=sha256:4a7badf5d18c7b80208b98d297ddd273691d9e9bd754a1fb11fe814b0c1aa867

Observation 56eb93fb-3c0f-401d-b7cd-674dd7a7498e · outbound

This paper cites GCBF+: A neural graph control barrier function framework for distributed safe multiagent control.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL GCBF+: A neural graph control barrier function framework for distributed safe multiagent control

Reference 87

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:35:50.055969Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-16T11:35:49.534940Z digest=sha256:311495278e9a2e10cac601329c56ef6b8132d88037b7c71d176f3064c66fd8f8

Observation 89e57848-5be3-4969-8f1a-b94fca74a625 · outbound

This paper cites MAMPS: Safe Multi-Agent Reinforcement Learning via Model Predictive Shielding.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL MAMPS: Safe Multi-Agent Reinforcement Learning via Model Predictive Shielding

Reference 88

Resolution
unresolved
no resolver link, observed 2026-08-16T11:35:49.539360Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:35:49.539360Z digest=sha256:a47d717cc223b126adaa1355add0e58112df1272c14aec418b2e897c233885c6

Observation cc1d6b07-680b-41d2-9377-4f86753bcc21 · outbound

This paper cites Model-free safe control for zero-violation reinforcement learning.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL Model-free safe control for zero-violation reinforcement learning

Reference 89

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:35:50.040156Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-16T11:35:49.544295Z digest=sha256:8e853a4a28b5459559246b8aa9eab49164d768b0a7e08c86354b2b19c0aec4c3

Observation 02e63a0a-c108-48df-95a8-786b32f80c81 · outbound

This paper cites Multi-agent first order con- strained optimization in policy space.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL Multi-agent first order con- strained optimization in policy space

Reference 90

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:35:50.023762Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-16T11:35:49.549368Z digest=sha256:ffbf620fe220e40ee1e63d7825f3dc37c9d42ab536e1c49e1b3dc4bc5a8b757a

Observation c23fb004-13ae-4b79-bebd-59a7e5ab4f6f · outbound

This paper cites Fast, on-line collision avoidance for dynamic vehicles using buffered voronoi cells.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL Fast, on-line collision avoidance for dynamic vehicles using buffered voronoi cells

Reference 91

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:35:50.006851Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-16T11:35:49.555561Z digest=sha256:fa073ac276dd9714c5d33bf0ac3a8bd16591350d1cfd57aad3e1f14f6f3290f7

Observation d8b0142c-0fc5-41ba-b1f1-3b1efdd930f5 · outbound

This paper cites Trajectory optimization for nonlinear multi-agent systems using decentralized learning model predictive control.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL Trajectory optimization for nonlinear multi-agent systems using decentralized learning model predictive control

Reference 92

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:35:49.990352Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-16T11:35:49.560635Z digest=sha256:03ac4d7e6295efa6383ac87eb52aa3e0a3f687cd55de17bb186e62166fbc82b7

Observation 26b1b56b-96f7-4307-bc1d-b75711f7a03f · outbound

This paper cites In other words, for a given z0, the value at the kth timestep is only a function of zk and xk instead of the z0 and the entire trajectory up to the kth timestep.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL In other words, for a given z0, the value at the kth timestep is only a function of zk and xk instead of the z0 and the entire trajectory up to the kth timestep

Reference 93

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:35:49.972993Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-16T11:35:49.565966Z digest=sha256:2b725925f8bf8a7e815a03c26a59c0c1b3b960993ed4e7f69e4281bb5cd70b26

Observation f0cdaa1e-0e66-4000-aae9-24dc084b845f · outbound

This paper cites an unresolved cited work.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL Unresolved cited work

Reference 94

Resolution
unresolved
raw_fallback, observed 2026-08-16T11:35:49.955745Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-16T11:35:49.570731Z digest=sha256:0fbf55a974377445c1b82067c5c435bf0fadee48f2a8dd08d933ed256ba74176

Observation 6cffaff4-cbbf-4eac-8f7f-8c27fe1a09b2 · outbound

This paper cites an unresolved cited work.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL Unresolved cited work

Reference 95

Resolution
unresolved
raw_fallback, observed 2026-08-16T11:35:49.938382Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-16T11:35:49.575826Z digest=sha256:277a81a2a8be701bcd7bbe516b237a07dba56f41559cb80fc1ed144193e4710d

Observation 53d0bdd6-665c-4045-a16d-ea6f6b03ab4d · outbound

This paper cites an unresolved cited work.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL Unresolved cited work

Reference 96

Resolution
unresolved
raw_fallback, observed 2026-08-16T11:35:49.922510Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-16T11:35:49.580801Z digest=sha256:11034a72edb62c14eaa4d16c73dce99d7f50c7f96b1731bef1ecbaffd7009955

Observation be1d4854-de22-48a7-876d-7e1571c4931d · outbound

This paper cites APPENDIX D ALGORITHM PSEUDOCODE We describe the centralized training process of Def-MARL in Algorithm 1 and the distributed execution process in Algorithm 2.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL APPENDIX D ALGORITHM PSEUDOCODE We describe the centralized training process of Def-MARL in Algorithm 1 and the distributed execution process in Algorithm 2

Reference 97

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:35:49.907043Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-16T11:35:49.585636Z digest=sha256:4cf50818908cb992b830289f7041fb793d29bb65e7da3a76e37d88b76074768e

Observation 99cdc52b-244d-4645-aa18-6dae73dd4be0 · outbound

This paper cites E⊆{ (i,j )|i∈V a,j ∈V} is the set of edges, denoting the information flow from a sender node j to a receiver agent i.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL E⊆{ (i,j )|i∈V a,j ∈V} is the set of edges, denoting the information flow from a sender node j to a receiver agent i

Reference 98

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:35:49.889638Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-16T11:35:49.590785Z digest=sha256:e3fe6e8862b06267ec9a894011cbbe0429c9e0efca3d78247a4a03da5d05e868

Observation 2884d2e1-7170-4d80-beb3-d044cd74f8c9 · outbound

This paper cites an unresolved cited work.

Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL Unresolved cited work

Reference 99

Resolution
unresolved
raw_fallback, observed 2026-08-16T11:35:49.870781Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-16T11:35:49.596258Z digest=sha256:88895fe96a41ab579fd2473d2f2990730e61b5dfa67124c4e67f7c3d56e8ebd5

Pith citing papers

Observation e8d1989e-45f5-44b0-bf8c-d9b029676de5 · inbound

Distributed Safety-Critical Control of Multi-Agent Systems with Time-Varying Communication Topologies cites this paper.

Distributed Safety-Critical Control of Multi-Agent Systems with Time-Varying Communication Topologies Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL

Reference 1

Resolution
verified exact
arxiv_id, observed 2026-05-13T22:58:23.933705Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-05-13T22:56:33.073457Z digest=sha256:fe3e6fde857be0e7517b18c4949bdb7259f11dbee4dce49a802f6981e44b70d3