Pith. sign in

Paper Citation Record · LEDGER

Nash Q-Network for Multi-Agent Cybersecurity Simulation

As of 23 August 2026, this Paper Citation Record lists 24 of 24 outbound references and 0 inbound Pith citation observations for arXiv:2509.00678.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2509.00678 v1

Coverage vector

measured 24 of 24 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-05T13:25:59.994792Z

measured 24 of 24 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-22T06:32:14.747728+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

24 of 24 outbound references displayed

  • verified exact0
  • verified fuzzy17
  • unresolved7
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation d0b33f9c-321f-4f19-a4ad-335bcc4d45c6 · outbound

This paper cites Mastering the game of go with deep neural networks and tree search,.

Nash Q-Network for Multi-Agent Cybersecurity Simulation Mastering the game of go with deep neural networks and tree search,

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T13:26:00.444103Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-05T13:25:59.883548Z digest=sha256:ab327e5c8b04bb87eea47675e949cccdbdef02f5e7ca5365838d71fbe249573c

Observation 58e3bcf5-7b2a-4f4a-b40d-cf899842ed07 · outbound

This paper cites Grandmaster level in starcraft ii using multi-agent reinforcement learning,.

Nash Q-Network for Multi-Agent Cybersecurity Simulation Grandmaster level in starcraft ii using multi-agent reinforcement learning,

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-05T13:25:59.890372Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T13:25:59.890372Z digest=sha256:60e0cc0e6459863c7605025eca0b92da1aee3febff49f2a544f69a1a50b74aec

Observation 654eb9cc-baa9-4716-9d03-a01492134ddb · outbound

This paper cites Nash q-learning for general-sum stochastic games,.

Nash Q-Network for Multi-Agent Cybersecurity Simulation Nash q-learning for general-sum stochastic games,

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T13:26:00.419153Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-05T13:25:59.895645Z digest=sha256:88387986b486b4b351371dca12ec290443167d5e5d87f0e74d22678c7b29e219

Observation 6739e7ce-1158-4286-841c-bb0de738dd03 · outbound

This paper cites CybORG: A Gym for the Development of Autonomous Cyber Agents.

Nash Q-Network for Multi-Agent Cybersecurity Simulation CybORG: A Gym for the Development of Autonomous Cyber Agents

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-05T13:25:59.900077Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T13:25:59.900077Z digest=sha256:70e18dee7eb9d90f7bb86321fd01e106e3c0e286213ec4081db79cfcd11160db

Observation cf402e31-f856-4e8f-8e9e-4ca5495a6105 · outbound

This paper cites Markov games as a framework for multi-agent reinforcement learning,.

Nash Q-Network for Multi-Agent Cybersecurity Simulation Markov games as a framework for multi-agent reinforcement learning,

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T13:26:00.403467Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-05T13:25:59.905139Z digest=sha256:18d1010e6cfe268285486c643e8d2f4f050c23e9f7e0d4f9cfe596437a5bf5e5

Observation a11558df-aaa1-4c02-a5b8-3504999e5596 · outbound

This paper cites Markov perfect equilibrium in a repeated principal-agent rela- tionship,.

Nash Q-Network for Multi-Agent Cybersecurity Simulation Markov perfect equilibrium in a repeated principal-agent rela- tionship,

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T13:26:00.387607Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-05T13:25:59.909628Z digest=sha256:6329d1de45afe51d1d9166dbac07178eda8db336b9f0a10209ae777e9035a128

Observation 660e8977-5268-4bde-9b3b-1a897a2692ab · outbound

This paper cites A survey of game theory as applied to network security,.

Nash Q-Network for Multi-Agent Cybersecurity Simulation A survey of game theory as applied to network security,

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T13:26:00.371273Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-05T13:25:59.914821Z digest=sha256:fa62a983dfc65f51cc60a14670ba4c6893f16cdbd5608fa3fe381ffa0bf41a7c

Observation 0e76ac92-a337-48c5-bae1-0506ba3f02b9 · outbound

This paper cites The complexity of computing a nash equilibrium,.

Nash Q-Network for Multi-Agent Cybersecurity Simulation The complexity of computing a nash equilibrium,

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-05T13:25:59.919127Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T13:25:59.919127Z digest=sha256:d82b0862b6c9307e77a22ac93ce9f7fe54fedd21c2439dd9d7a02e82c1ec10b5

Observation 47ba5eb8-7431-4792-a86f-2b4bd112a522 · outbound

This paper cites A comprehensive survey of multiagent reinforce- ment learning,.

Nash Q-Network for Multi-Agent Cybersecurity Simulation A comprehensive survey of multiagent reinforce- ment learning,

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T13:26:00.344338Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-05T13:25:59.927226Z digest=sha256:014819e650647f2b4805344d2d319f15f5ae74930d82733f378d05fd446f5e65

Observation 791f92c5-7d29-4357-bcfe-32805ab67298 · outbound

This paper cites A deep learning-based multi-agent system for intrusion detection,.

Nash Q-Network for Multi-Agent Cybersecurity Simulation A deep learning-based multi-agent system for intrusion detection,

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T13:26:00.328910Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-05T13:25:59.931216Z digest=sha256:b0c9f60157024e8863c3f9db03d620afa91abe3feafda1be5800e669888b6c90

Observation 1536871c-4b62-4224-b043-26e0a91e9ca7 · outbound

This paper cites Deep reinforcement learning for adaptive cyber defense in network security,.

Nash Q-Network for Multi-Agent Cybersecurity Simulation Deep reinforcement learning for adaptive cyber defense in network security,

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T13:26:00.312338Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-05T13:25:59.936139Z digest=sha256:eb13f519a213852c02390f4368f1789f64960434cfb078a3a817e1bfece7206b

Observation 268408f9-a9c1-48cf-b09c-44dcfedf9ec3 · outbound

This paper cites Reinforcement learning for efficient network penetration testing,.

Nash Q-Network for Multi-Agent Cybersecurity Simulation Reinforcement learning for efficient network penetration testing,

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T13:26:00.295366Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-05T13:25:59.940401Z digest=sha256:948c32ed1b390ba1e4a5c4719c0446741abe38c9fd3a707e8b93a0efbf8602dd

Observation bc8725fc-5fec-4e80-9946-8621ad8eff54 · outbound

This paper cites Combining deep reinforcement learning and search for imperfect-information games,.

Nash Q-Network for Multi-Agent Cybersecurity Simulation Combining deep reinforcement learning and search for imperfect-information games,

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T13:26:00.274921Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-05T13:25:59.945508Z digest=sha256:83a269afa798982229032632e0770f4d18eb3648159e1db375e0f717e730bda6

Observation c3f39add-fc6a-4cd0-a7fd-6abbe4244ead · outbound

This paper cites Safe Multi-Agent Reinforcement Learning with Convergence to Generalized Nash Equilibrium.

Nash Q-Network for Multi-Agent Cybersecurity Simulation Safe Multi-Agent Reinforcement Learning with Convergence to Generalized Nash Equilibrium

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-05T13:25:59.949703Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T13:25:59.949703Z digest=sha256:55beabc41a0f9fa3d214482e65ee7f8c1c71efc780b57aa3c0419260818bc0bb

Observation c50cc2f2-d6a6-4d23-9237-2f8437e50846 · outbound

This paper cites A comprehensive systematic literature review on intrusion detection systems,.

Nash Q-Network for Multi-Agent Cybersecurity Simulation A comprehensive systematic literature review on intrusion detection systems,

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T13:26:00.257456Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-05T13:25:59.954379Z digest=sha256:d6c9e7c0322fb6b5f5a6e38b3bcb3d1a2275448838ebd1f8cb4848ed4d0594f5

Observation 3cea5c4b-63d2-4098-8b7e-415fc20a21d2 · outbound

This paper cites Correlated q-learning,.

Nash Q-Network for Multi-Agent Cybersecurity Simulation Correlated q-learning,

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T13:26:00.240980Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-05T13:25:59.959210Z digest=sha256:94ccfe631500c00a4f69ffc172695cce09b90da44d5b1b5582e4cb024f9f8da6

Observation 83fe8b63-e8fd-4d44-8fdd-c4361c9d43d6 · outbound

This paper cites Alpcan and T.

Nash Q-Network for Multi-Agent Cybersecurity Simulation Alpcan and T

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T13:26:00.222553Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-05T13:25:59.963344Z digest=sha256:fbe9ccf9827505e4ed103e61f38689ecb30f3bd27812e9680d7c5af14d929379

Observation bac529b3-d118-43e0-a6da-8002aa74cc69 · outbound

This paper cites Deep q-learning for nash equilibria: Nash-dqn,.

Nash Q-Network for Multi-Agent Cybersecurity Simulation Deep q-learning for nash equilibria: Nash-dqn,

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T13:26:00.201116Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-05T13:25:59.967510Z digest=sha256:281fe3425a79929a328b42ee8cca08e97cbfc9fd8347ebfaf6fd24696d22f8c0

Observation 619f91d3-2eb0-4a11-8fcb-73c5d162bd0c · outbound

This paper cites Equilibrium points of bimatrix games,.

Nash Q-Network for Multi-Agent Cybersecurity Simulation Equilibrium points of bimatrix games,

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-05T13:25:59.971718Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T13:25:59.971718Z digest=sha256:01baf25e2005066436d82a1c0aedecec1fe5c9639e8001ff414230a90505b2a7

Observation 50a2a5f4-afb0-4d2e-b95d-8162a1f50534 · outbound

This paper cites Ray: A distributed framework for emerging{AI} applications,.

Nash Q-Network for Multi-Agent Cybersecurity Simulation Ray: A distributed framework for emerging{AI} applications,

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T13:26:00.164551Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-05T13:25:59.976476Z digest=sha256:66d70a1b0d67e17e5d7e2a0e79553352b4b35f99e71e9c2ac737eadabe6d086d

Observation 74578041-7280-4e3b-ab31-5f30c4d06c35 · outbound

This paper cites Adam: A Method for Stochastic Optimization.

Nash Q-Network for Multi-Agent Cybersecurity Simulation Adam: A Method for Stochastic Optimization

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-05T13:25:59.981069Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T13:25:59.981069Z digest=sha256:10b96d043d2648622d61386cb65c6f5730ea59f62d7cd3397302878196720c0b

Observation d4e2f42a-cbbf-4ef5-a294-dbad97247885 · outbound

This paper cites On Autonomous Agents in a Cyber Defence Environment.

Nash Q-Network for Multi-Agent Cybersecurity Simulation On Autonomous Agents in a Cyber Defence Environment

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-05T13:25:59.985731Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T13:25:59.985731Z digest=sha256:230cc3edd691aea66e485083186d3f9a51547af6ceff83e76ada3910bad37b12

Observation b93b0a16-c381-438c-b80f-1aa0ad38d38f · outbound

This paper cites Autonomous network defence using reinforcement learning,.

Nash Q-Network for Multi-Agent Cybersecurity Simulation Autonomous network defence using reinforcement learning,

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T13:26:00.147556Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-05T13:25:59.990406Z digest=sha256:dde1ed4f3bf5bf5e26583b67f93d58671e471b6ab7cf557018ebcd8457a15da8

Observation fffac843-3bda-41fa-83e2-23177bb0b5d4 · outbound

This paper cites Learning to communicate in multi-agent reinforcement learning for au- tonomous cyber defence,.

Nash Q-Network for Multi-Agent Cybersecurity Simulation Learning to communicate in multi-agent reinforcement learning for au- tonomous cyber defence,

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T13:26:00.127826Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-05T13:25:59.994792Z digest=sha256:80a23dfb22ea64c8adb3bca14d6172e6f750ebc380f3381574e2ca3b398c69e2

Pith citing papers

No inbound Pith citation observations are available.