Pith. sign in

Paper Citation Record · LEDGER

A Lyapunov Drift-Plus-Penalty Method Tailored for Reinforcement Learning with Queue Stability

As of 9 August 2026, this Paper Citation Record lists 29 of 29 outbound references and 1 inbound Pith citation observation for arXiv:2506.04291.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.04291 v1

Coverage vector

measured 29 of 29 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T11:01:07.763959Z

measured 30 of 30 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-06-27T19:06:08.261216Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-02T22:17:25.890997Z

Reference resolution

29 of 29 outbound references displayed

  • verified exact0
  • verified fuzzy28
  • unresolved1
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 3158a2da-e0a7-439d-9921-3e4b1910c18b · outbound

This paper cites Distributed Lyapunov drift-plus-penalty routing for wifi mesh networks with adaptive penalty weight,.

A Lyapunov Drift-Plus-Penalty Method Tailored for Reinforcement Learning with Queue Stability Distributed Lyapunov drift-plus-penalty routing for wifi mesh networks with adaptive penalty weight,

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:01:12.117028Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T11:01:06.323643Z digest=sha256:b7fc6ff41dd455669d0c914b6eb0fd1f4be856b0aa0643ac9c1fb6bf5d76a7dd

Observation 7f4d71af-add1-4464-a895-de3cbb72019a · outbound

This paper cites Lyapunov optimized resource management for multiuser mobile video streaming,.

A Lyapunov Drift-Plus-Penalty Method Tailored for Reinforcement Learning with Queue Stability Lyapunov optimized resource management for multiuser mobile video streaming,

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:01:11.943084Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T11:01:06.348003Z digest=sha256:81b0a106a528d4dab8027051013e84fdc502558b03786d6bd5928c8c08ba9942

Observation 02a61fff-48a5-4775-be65-be9b6cd01d9d · outbound

This paper cites Lyapunov optimization framework for 5G mobile nodes with multi-homing,.

A Lyapunov Drift-Plus-Penalty Method Tailored for Reinforcement Learning with Queue Stability Lyapunov optimization framework for 5G mobile nodes with multi-homing,

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:01:11.765546Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T11:01:06.374650Z digest=sha256:f2cf0aff6360ab012c8f3f870fc90a60479e2ae06d52071ec44c4835c7d4b321

Observation dd432814-721d-4b28-b977-35b93c37a764 · outbound

This paper cites On the convergence time of the drift-plus- penalty algorithm for strongly convex programs,.

A Lyapunov Drift-Plus-Penalty Method Tailored for Reinforcement Learning with Queue Stability On the convergence time of the drift-plus- penalty algorithm for strongly convex programs,

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:01:11.607650Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T11:01:06.405104Z digest=sha256:836f4d5c410e25ce114db6ffb1aa029d9279b4acc54f5e9e3a35bf5669cb3276

Observation d931f671-c29a-4ced-beb1-67eb10e8267f · outbound

This paper cites Dynamic resource allocation in metro elastic optical networks using Lyapunov drift optimization,.

A Lyapunov Drift-Plus-Penalty Method Tailored for Reinforcement Learning with Queue Stability Dynamic resource allocation in metro elastic optical networks using Lyapunov drift optimization,

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:01:11.443356Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T11:01:06.424819Z digest=sha256:2ed2b51932da0d4f644ccdfb1790f17fd1413b6406834728001a10863ef45258

Observation a15c32da-1c38-4c53-bd46-0d05946ae1a2 · outbound

This paper cites Neely,Stochastic network optimization with application to commu- nication and queueing systems.

A Lyapunov Drift-Plus-Penalty Method Tailored for Reinforcement Learning with Queue Stability Neely,Stochastic network optimization with application to commu- nication and queueing systems

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:01:11.266248Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T11:01:06.442024Z digest=sha256:f1f5dbbe98053b568f2fa7dbc6d0f001e2cb41e1370399edeb71435582a8c340

Observation e341be66-e734-4a70-b5c5-8dc7efcb74a1 · outbound

This paper cites On the optimality of greedy policies in dynamic matching,.

A Lyapunov Drift-Plus-Penalty Method Tailored for Reinforcement Learning with Queue Stability On the optimality of greedy policies in dynamic matching,

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:01:11.112886Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T11:01:06.460847Z digest=sha256:f8e8b378fcae803515582ec6459763631594aa3d5c9c6f09b5efaaeb0e1739bd

Observation 94f6d505-d3f3-4377-8fc6-99c74b34faa3 · outbound

This paper cites Deep learning for channel tracking in IRS-assisted UA V communication systems,.

A Lyapunov Drift-Plus-Penalty Method Tailored for Reinforcement Learning with Queue Stability Deep learning for channel tracking in IRS-assisted UA V communication systems,

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:01:10.936179Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T11:01:06.476628Z digest=sha256:981ba7dbc531301c7a9d3910dd650196e7eedb45d7bd5a58c5b2a0ad19a25ba0

Observation a3f58c3a-a934-4824-aa8a-a4abb3091187 · outbound

This paper cites Lyapunov-guided deep reinforcement learning for stable online computation offloading in mobile-edge computing networks,.

A Lyapunov Drift-Plus-Penalty Method Tailored for Reinforcement Learning with Queue Stability Lyapunov-guided deep reinforcement learning for stable online computation offloading in mobile-edge computing networks,

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:01:10.770761Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T11:01:06.495553Z digest=sha256:4bc277d87afcebc2e6c40c0e1793009e23a92441d981923d34a0eaebda0b7766

Observation c407d6cb-4496-444b-9b3a-e3020149f66b · outbound

This paper cites Multi-agent deep reinforcement learning for computation offloading and interference coordination in small cell networks,.

A Lyapunov Drift-Plus-Penalty Method Tailored for Reinforcement Learning with Queue Stability Multi-agent deep reinforcement learning for computation offloading and interference coordination in small cell networks,

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:01:10.589231Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T11:01:06.514484Z digest=sha256:1ecedaadab19b56c465ec55dab7a358448cb7013be6ae8e1715134170f8d9a85

Observation ce9eb58e-9107-4cd8-b330-6889f4045a96 · outbound

This paper cites Neely,Stochastic network optimization with application to commu- nication and queueing systems.

A Lyapunov Drift-Plus-Penalty Method Tailored for Reinforcement Learning with Queue Stability Neely,Stochastic network optimization with application to commu- nication and queueing systems

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-07T11:01:06.535745Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:01:06.535745Z digest=sha256:4fe5de00adb556d9a67a25ebfde25e2422ec70e59caf6296cb13d018145a40fe

Observation af631270-0a47-4902-9fc3-13dd05401dda · outbound

This paper cites When Lyapunov drift meets DRL: Energy efficient resource allocation for IoT data collecting,.

A Lyapunov Drift-Plus-Penalty Method Tailored for Reinforcement Learning with Queue Stability When Lyapunov drift meets DRL: Energy efficient resource allocation for IoT data collecting,

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:01:10.363765Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T11:01:06.577748Z digest=sha256:3fcdd04f47ebbfdcfe16b5f2a7c24473b3845ecd597fafc61f81b06d9abf9870

Observation a30643d8-fc0b-4926-a998-bd7906edca31 · outbound

This paper cites Lyapunov-guided resource allocation and task scheduling for edge computing cognitive radio networks via deep reinforcement learning,.

A Lyapunov Drift-Plus-Penalty Method Tailored for Reinforcement Learning with Queue Stability Lyapunov-guided resource allocation and task scheduling for edge computing cognitive radio networks via deep reinforcement learning,

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:01:10.166582Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T11:01:06.602334Z digest=sha256:4142631e307214f557ff7e9e0f17a750b157450a6c8cb359e2afac8fe9cb6acc

Observation 6472dff7-fc97-4787-85b1-781bfd2dc161 · outbound

This paper cites Lyapunov drift-plus-penalty optimization for queues with finite capacity,.

A Lyapunov Drift-Plus-Penalty Method Tailored for Reinforcement Learning with Queue Stability Lyapunov drift-plus-penalty optimization for queues with finite capacity,

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:01:09.989110Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T11:01:06.631966Z digest=sha256:b37b13e87ebc0de4adfe0e10edc7679f09295f281869afdf66ea65cfaa5a8d1e

Observation 5903d4ec-2b3f-49fd-b648-e156eefe3754 · outbound

This paper cites Joint task offloading and resource allocation for energy-constrained mobile edge computing,.

A Lyapunov Drift-Plus-Penalty Method Tailored for Reinforcement Learning with Queue Stability Joint task offloading and resource allocation for energy-constrained mobile edge computing,

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:01:09.840118Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T11:01:06.664035Z digest=sha256:77713555d4d973a1c2f88bf2e366b16730d7c3c11940ec0256f4440a1f13499c

Observation 83bc2677-7f8a-44d1-830f-abc29cba5d69 · outbound

This paper cites UA V-assisted task offloading in vehicular edge computing networks,.

A Lyapunov Drift-Plus-Penalty Method Tailored for Reinforcement Learning with Queue Stability UA V-assisted task offloading in vehicular edge computing networks,

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:01:09.691597Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T11:01:06.697182Z digest=sha256:cb6eff0de9cb5e505a9e507b26541ee0fe8e224a00cb07c74f8e48804a54d268

Observation 711865ac-568a-4862-85e1-058120b0d140 · outbound

This paper cites Energy-efficient federated edge learning with streaming data: A lyapunov optimization approach,.

A Lyapunov Drift-Plus-Penalty Method Tailored for Reinforcement Learning with Queue Stability Energy-efficient federated edge learning with streaming data: A lyapunov optimization approach,

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:01:09.519959Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T11:01:06.727727Z digest=sha256:1d0ec56ab74f5f3623d6b8f619a4f6eb0405951344aeb71d05132e571702fe4c

Observation 6c985a2f-3c2e-4d63-a0c7-4630dc8e8a13 · outbound

This paper cites Profit maxi- mization of independent task offloading in MEC-enabled 5G internet of vehicles,.

A Lyapunov Drift-Plus-Penalty Method Tailored for Reinforcement Learning with Queue Stability Profit maxi- mization of independent task offloading in MEC-enabled 5G internet of vehicles,

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:01:09.390067Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T11:01:06.776689Z digest=sha256:eb17ccc8c5e45e32c7edde41789f74ca1268ecb63ce207128def97105d9dd195

Observation 22f6739b-4a13-470a-a69c-8d94dcd869b0 · outbound

This paper cites An online rein- forcement learning-based energy management strategy for microgrids with centralized control,.

A Lyapunov Drift-Plus-Penalty Method Tailored for Reinforcement Learning with Queue Stability An online rein- forcement learning-based energy management strategy for microgrids with centralized control,

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:01:09.250458Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T11:01:06.878338Z digest=sha256:68fa2865bc7267fc9c20b560142f1e3d9b70f02d618ea128856d4a67d1fac58c

Observation 412d03e4-1991-428c-a500-bee564655913 · outbound

This paper cites Offline meta- reinforcement learning for active pantograph control in high-speed railways,.

A Lyapunov Drift-Plus-Penalty Method Tailored for Reinforcement Learning with Queue Stability Offline meta- reinforcement learning for active pantograph control in high-speed railways,

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:01:09.123785Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T11:01:06.982304Z digest=sha256:a686794cf30f954b91b55cbf02401f7f1be7098abb83215c2d80745dde7e7fa3

Observation 6d804eaf-b921-4610-933c-6a4a2169b5cb · outbound

This paper cites Survey on large language model-enhanced reinforcement learning: Concept, taxonomy, and methods,.

A Lyapunov Drift-Plus-Penalty Method Tailored for Reinforcement Learning with Queue Stability Survey on large language model-enhanced reinforcement learning: Concept, taxonomy, and methods,

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:01:09.002053Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T11:01:07.065216Z digest=sha256:3957405b3d767536ab1be1a3805460e8d15c2567d607222f9e9374d047cdace6

Observation 53137d09-bee3-4958-aa30-8210f825c9b8 · outbound

This paper cites Developments in image processing using deep learning and reinforcement learning,.

A Lyapunov Drift-Plus-Penalty Method Tailored for Reinforcement Learning with Queue Stability Developments in image processing using deep learning and reinforcement learning,

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:01:08.870249Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T11:01:07.144477Z digest=sha256:e63fdc88cf25360379a1c76c85ea5561d258ab253cc55235fb9fbc60e9b50492

Observation 680bba6e-828f-4048-8281-57e8799b8362 · outbound

This paper cites Maximum entropy reinforcement learn- ing in two-player perfect information games,.

A Lyapunov Drift-Plus-Penalty Method Tailored for Reinforcement Learning with Queue Stability Maximum entropy reinforcement learn- ing in two-player perfect information games,

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:01:08.733731Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T11:01:07.262197Z digest=sha256:33d8360cfa9742057624e8848992a7c03d3b1d30da4daa920c75a6a28e35c9c1

Observation 267f8b6f-1ad6-417c-a349-9cf0d5ee5c7f · outbound

This paper cites Deep reinforcement learning for stochastic computation offloading in digital twin networks,.

A Lyapunov Drift-Plus-Penalty Method Tailored for Reinforcement Learning with Queue Stability Deep reinforcement learning for stochastic computation offloading in digital twin networks,

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:01:08.606839Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T11:01:07.339429Z digest=sha256:7f014c548eb15f1384294675e6c183022fff8d1d2df336e630ebd60413d8cfcb

Observation 7544f824-5bd4-4ac1-b54f-30113757517d · outbound

This paper cites Learning-based context-aware resource allocation for edge-computing-empowered industrial IoT,.

A Lyapunov Drift-Plus-Penalty Method Tailored for Reinforcement Learning with Queue Stability Learning-based context-aware resource allocation for edge-computing-empowered industrial IoT,

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:01:08.471819Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T11:01:07.462947Z digest=sha256:5576e0dd8f1eb1c00ddb221968dca2ed5679c641c3aca345eb755625144f7f68

Observation ce2291f5-3dab-4953-b7f9-7660edd81e0c · outbound

This paper cites Predictable wireless networked scheduling for bridging hybrid time-sensitive and real-time services,.

A Lyapunov Drift-Plus-Penalty Method Tailored for Reinforcement Learning with Queue Stability Predictable wireless networked scheduling for bridging hybrid time-sensitive and real-time services,

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:01:08.338994Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T11:01:07.572313Z digest=sha256:0204a4e44d15f6f615abb3bdcc79a97bf5b5f627153343c36a2cf72a5dd0de81

Observation 3b55bda4-7f0d-425c-b83a-9ed6ce38abb7 · outbound

This paper cites Lyapunov optimization based mobile edge computing for internet of vehicles systems,.

A Lyapunov Drift-Plus-Penalty Method Tailored for Reinforcement Learning with Queue Stability Lyapunov optimization based mobile edge computing for internet of vehicles systems,

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:01:08.207066Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T11:01:07.597884Z digest=sha256:6de5a2394ac2a7fb9b9511b7e33cbc63fa8846d7ec04380e91a11c243e9de97a

Observation 596603da-452c-40b0-a9bf-00467693723a · outbound

This paper cites Enhancing fog computing through intelligent reflecting surface assistance: A Lyapunov driven reinforcement learning approach,.

A Lyapunov Drift-Plus-Penalty Method Tailored for Reinforcement Learning with Queue Stability Enhancing fog computing through intelligent reflecting surface assistance: A Lyapunov driven reinforcement learning approach,

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:01:08.078244Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T11:01:07.673486Z digest=sha256:1e566b6fdd5cc3ad7e0a39ea6104e4aaed5a1d6988c2c5deb47d31d1bae73cf4

Observation 915105e3-3fdf-4f08-b4ff-63306c3414d7 · outbound

This paper cites Energy-latency aware intelligent reflecting surface aided multi-cell mobile edge computing,.

A Lyapunov Drift-Plus-Penalty Method Tailored for Reinforcement Learning with Queue Stability Energy-latency aware intelligent reflecting surface aided multi-cell mobile edge computing,

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:01:07.902668Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T11:01:07.763959Z digest=sha256:b2c0783f420e6eb6db87eeb3155967e9e091e60aea23b4e45db2be69b99b1b55

Pith citing papers

Observation f1b1b228-a3fe-492e-a223-d9211459a611 · inbound

Theoretical Foundations of Continual Learning via Drift-Plus-Penalty cites this paper.

Theoretical Foundations of Continual Learning via Drift-Plus-Penalty A Lyapunov Drift-Plus-Penalty Method Tailored for Reinforcement Learning with Queue Stability

Reference 9

Resolution
verified exact
arxiv_id, observed 2026-07-02T22:17:25.892783Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-27T19:06:08.261216Z digest=sha256:42643825927fd53a7f3e09f5077a827c19bcbd18248be50be5cf51f7a3639d75