Pith. sign in

Paper Citation Record · LEDGER

Towards Fault Tolerance in Multi-Agent Reinforcement Learning

As of 14 August 2026, this Paper Citation Record lists 54 of 54 outbound references and 1 inbound Pith citation observation for arXiv:2412.00534.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2412.00534 v1

Coverage vector

measured 54 of 54 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-12T05:22:09.722683Z

measured 55 of 55 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-14T06:32:32.682623+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-11T19:28:52.274750Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-11T19:28:52.638301Z

Reference resolution

54 of 54 outbound references displayed

  • verified exact1
  • verified fuzzy41
  • unresolved12
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 0ef39b6a-8f2d-4f3d-bc95-80bef705377f · outbound

This paper cites Deep reinforcement learning for autonomous driving: A survey,.

Towards Fault Tolerance in Multi-Agent Reinforcement Learning Deep reinforcement learning for autonomous driving: A survey,

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T05:22:10.544565Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T05:22:09.466136Z digest=sha256:20e20cca723382cb8123fcd2f95cb56e206ea9c45d7377f0629d5885c8597ae4

Observation 02149e4c-a161-4e55-b481-7c18aa4cb1ba · outbound

This paper cites Distributed multi-vehicle task assignment and motion planning in dense environments,.

Towards Fault Tolerance in Multi-Agent Reinforcement Learning Distributed multi-vehicle task assignment and motion planning in dense environments,

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T05:22:10.529434Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T05:22:09.472272Z digest=sha256:87f25b06df5fa6a61c6f1b3d39a551dbd1c27c3f0140800e0c4a750cc7b3e32d

Observation 8b42275d-fba1-4b60-b9fc-5787d72d0022 · outbound

This paper cites A survey on multi- agent reinforcement learning applications in the internet of vehicles,.

Towards Fault Tolerance in Multi-Agent Reinforcement Learning A survey on multi- agent reinforcement learning applications in the internet of vehicles,

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T05:22:10.514197Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T05:22:09.477116Z digest=sha256:3ffd3b84cb71035c66c743088a2f6cc90a85473c17e3f29dfcc01656908500eb

Observation 8cb1b92e-4780-4df8-b52e-499139071318 · outbound

This paper cites Theory and experiment on formation-containment control of multiple multirotor unmanned aerial vehicle systems,.

Towards Fault Tolerance in Multi-Agent Reinforcement Learning Theory and experiment on formation-containment control of multiple multirotor unmanned aerial vehicle systems,

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T05:22:10.498862Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T05:22:09.482163Z digest=sha256:3fd1925f36ae7d58fa4fee9687f13c008c36940139a951945b2bce0563ec9cfd

Observation f0306bee-2680-44c6-b6b5-6a3e1a81ed80 · outbound

This paper cites Cooperative internet of uavs: Distributed trajectory design by multi-agent deep reinforcement learning,.

Towards Fault Tolerance in Multi-Agent Reinforcement Learning Cooperative internet of uavs: Distributed trajectory design by multi-agent deep reinforcement learning,

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T05:22:10.483609Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T05:22:09.487265Z digest=sha256:4c7ebb7035630576f9f93fb2a5462e45355693bb751c6c020732a80cf938227c

Observation 5604ac6e-8397-4079-b61e-59edd446b198 · outbound

This paper cites Heterogeneous multi-robot reinforcement learning,.

Towards Fault Tolerance in Multi-Agent Reinforcement Learning Heterogeneous multi-robot reinforcement learning,

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-12T05:22:09.492232Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T05:22:09.492232Z digest=sha256:7f8e2030f094cef1b2036b8228765a48f22df48054b8d87a50877c7aa2d5dca7

Observation fd591eb3-54e0-4f71-b9f6-45b77ed9788b · outbound

This paper cites On the effects of communication failures in a multi-agent consensus network,.

Towards Fault Tolerance in Multi-Agent Reinforcement Learning On the effects of communication failures in a multi-agent consensus network,

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T05:22:10.459235Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T05:22:09.497616Z digest=sha256:5e590d8620fd28a10b0d3b67a9057eca870fd4c75efd76f68f661fe28e86019a

Observation a10027c5-14bf-48ec-bc2d-5f621d29fc50 · outbound

This paper cites The impact of agent definitions and interactions on multiagent learning for coordination in traffic management domains,.

Towards Fault Tolerance in Multi-Agent Reinforcement Learning The impact of agent definitions and interactions on multiagent learning for coordination in traffic management domains,

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T05:22:10.444134Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T05:22:09.502419Z digest=sha256:3ed7f9bc527cdef7eb013e27ae206c2cf536a8f1a3889ea5e731157c22940b03

Observation 42d4a4d0-ae76-404d-babf-db761ca0b3e6 · outbound

This paper cites Fault-tolerant cooperative control of multiagent systems: A survey of trends and methodologies,.

Towards Fault Tolerance in Multi-Agent Reinforcement Learning Fault-tolerant cooperative control of multiagent systems: A survey of trends and methodologies,

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T05:22:10.428546Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T05:22:09.507161Z digest=sha256:966b936d20c1f270331e1363f8bfef62a79dfd287cf2e3b6a2508356dd53f13c

Observation aaaee778-8470-425c-9848-1389f31a27fa · outbound

This paper cites Fault-tolerant cooperative control of multiagent systems: A survey of trends and methodologies,.

Towards Fault Tolerance in Multi-Agent Reinforcement Learning Fault-tolerant cooperative control of multiagent systems: A survey of trends and methodologies,

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T05:22:10.413639Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T05:22:09.512052Z digest=sha256:1f93ad2b91188d25481b89ba5982097bc2ab9d6511bbe7f99ad4e78089ad180f

Observation d88154f3-590d-4c30-9103-2a4680edbef8 · outbound

This paper cites Fault-tolerant consensus of leader–following multi-agent systems with jointly connected topologies,.

Towards Fault Tolerance in Multi-Agent Reinforcement Learning Fault-tolerant consensus of leader–following multi-agent systems with jointly connected topologies,

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T05:22:10.398213Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T05:22:09.516793Z digest=sha256:7dac26a19fb021dfacd5d454bb82ea3246f33c10a679c5b1c7d901562f02df93

Observation f8c816c0-6ecb-4262-8606-061d5dc73661 · outbound

This paper cites Fault-tolerant formation for multi-uav via improved artificial potential field method,.

Towards Fault Tolerance in Multi-Agent Reinforcement Learning Fault-tolerant formation for multi-uav via improved artificial potential field method,

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T05:22:10.383273Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T05:22:09.521317Z digest=sha256:ae40e4ba0be3b6fa60ef94c0912642d12f605c4ef5941ce7a9cdbfb3836b8784

Observation 258fada4-c201-43a9-8e77-01c1bad04081 · outbound

This paper cites Policy gradient methods for reinforcement learning with function approximation,.

Towards Fault Tolerance in Multi-Agent Reinforcement Learning Policy gradient methods for reinforcement learning with function approximation,

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T05:22:10.367941Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T05:22:09.526089Z digest=sha256:1cb55145b6096a877befd01a05f02e456b9036f574846af650fd5ec1241ad95c

Observation 1334f93f-0e8b-4e12-9a4b-1aa40f17bb29 · outbound

This paper cites Multi-agent reinforcement learning with decentral- ized distribution correction,.

Towards Fault Tolerance in Multi-Agent Reinforcement Learning Multi-agent reinforcement learning with decentral- ized distribution correction,

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T05:22:10.353476Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T05:22:09.530793Z digest=sha256:bf305991ce845ecddee9e4327a33bee32222ca6f5ff74e50721bffd6b67fa04f

Observation 6171fcb6-9ed1-494d-9225-7a9c5e2d6209 · outbound

This paper cites Multi-agent actor-critic for mixed cooperative-competitive environments,.

Towards Fault Tolerance in Multi-Agent Reinforcement Learning Multi-agent actor-critic for mixed cooperative-competitive environments,

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T05:22:10.337665Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T05:22:09.535322Z digest=sha256:5de434f7673070dfd7c33c3a7ef40b9749b25bd5defe2d5a626574ad79ca711f

Observation 9d2b276e-728d-488a-8681-3dc6d323c42e · outbound

This paper cites The surprising effectiveness of ppo in cooperative multi-agent games,.

Towards Fault Tolerance in Multi-Agent Reinforcement Learning The surprising effectiveness of ppo in cooperative multi-agent games,

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T05:22:10.321728Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T05:22:09.539877Z digest=sha256:3c05cb3f7a8f66eaf329a0c9fdb08744342de1dccd1fa1eccdb74840a25685b7

Observation 14218bdb-2ea6-48bb-9e11-be1778733a3c · outbound

This paper cites Prioritized experience replay,.

Towards Fault Tolerance in Multi-Agent Reinforcement Learning Prioritized experience replay,

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T05:22:10.306904Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T05:22:09.544780Z digest=sha256:72a2eb1393618ef6e2e82645d85eddd70cdd65999954fcacb0d4acc4be0d4e0f

Observation 9ad8e1a3-40cf-42ff-b466-9bee99a61da5 · outbound

This paper cites Towards a fault-tolerant multi-agent system architecture,.

Towards Fault Tolerance in Multi-Agent Reinforcement Learning Towards a fault-tolerant multi-agent system architecture,

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T05:22:10.291970Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T05:22:09.549183Z digest=sha256:8e8baf18b7dfc646456c03e2ddf489928ba48102d21119959c19f9d2b6fe9d19

Observation 4e52e42a-8e2f-4624-9559-b0d513c7ee8a · outbound

This paper cites A survey on fault tolerant multi agent system,.

Towards Fault Tolerance in Multi-Agent Reinforcement Learning A survey on fault tolerant multi agent system,

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T05:22:10.276061Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T05:22:09.553810Z digest=sha256:55b4de5ae75bef54a3066f4029cdcb8d80dc43466d498dadac8493fe0de39bc1

Observation c2fb3118-444a-4c1d-87aa-8a0c187705c7 · outbound

This paper cites Adaptive fault-tolerant boundary control of an autonomous aerial refueling hose system with prescribed constraints,.

Towards Fault Tolerance in Multi-Agent Reinforcement Learning Adaptive fault-tolerant boundary control of an autonomous aerial refueling hose system with prescribed constraints,

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T05:22:10.260737Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T05:22:09.558575Z digest=sha256:590e84a890183c2b52701acf694b116f03cbf3f1829449e9a0c0239aa153517f

Observation 3f25025c-149c-4ea3-b7ed-07163fbf7e69 · outbound

This paper cites A goa-based fault-tolerant trajectory tracking control for an underwater vehicle of multi-thruster system without actuator saturation,.

Towards Fault Tolerance in Multi-Agent Reinforcement Learning A goa-based fault-tolerant trajectory tracking control for an underwater vehicle of multi-thruster system without actuator saturation,

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T05:22:10.244285Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T05:22:09.563422Z digest=sha256:fc57fcd2f419e6ac6a2b0dbf6c42d838b9208c9bd5153060780383b3074914d8

Observation 3bb7a928-8527-43a1-b731-49130c1e179a · outbound

This paper cites Fault-tolerant cooperative driving at signal-free intersections,.

Towards Fault Tolerance in Multi-Agent Reinforcement Learning Fault-tolerant cooperative driving at signal-free intersections,

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T05:22:10.229370Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T05:22:09.568080Z digest=sha256:8c551adadb73f732a128d1158e581c16e83767f102e097fd297bd3c169feac13

Observation 3abc5b53-d14b-4aac-a2ab-d4ff4f9f6519 · outbound

This paper cites Fault-tolerant cooperative control design of multiple wheeled mobile robots,.

Towards Fault Tolerance in Multi-Agent Reinforcement Learning Fault-tolerant cooperative control design of multiple wheeled mobile robots,

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T05:22:10.214535Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T05:22:09.572613Z digest=sha256:e28ae9b418876e2f090d8e2534ad69202fa97f86bcca802c79b9f9f1b582fbc8

Observation afa9543b-2763-4f44-9669-63ff2735f6ce · outbound

This paper cites Robust multi- agent reinforcement learning via minimax deep deterministic policy gradient,.

Towards Fault Tolerance in Multi-Agent Reinforcement Learning Robust multi- agent reinforcement learning via minimax deep deterministic policy gradient,

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T05:22:10.199425Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T05:22:09.576878Z digest=sha256:d3c6cb9d33254a8484e3801a3942d6c2b4ef5f2b35f6ba4a119d4e16eb611246

Observation bd5bb53a-38a5-4ccd-a9eb-65007ce2c0ec · outbound

This paper cites Byzantine Robust Cooperative Multi-Agent Reinforcement Learning as a Bayesian Game.

Towards Fault Tolerance in Multi-Agent Reinforcement Learning Byzantine Robust Cooperative Multi-Agent Reinforcement Learning as a Bayesian Game

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-12T05:22:09.581417Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T05:22:09.581417Z digest=sha256:1e9f9606185a4976732b8fd2e994bb2bd047a9a377cee0133b14224bc238312f

Observation 0ef13575-f180-404c-af0e-ae2b6765cf49 · outbound

This paper cites QMIX: Monotonic value function factorisation for deep multi-agent reinforcement learning,.

Towards Fault Tolerance in Multi-Agent Reinforcement Learning QMIX: Monotonic value function factorisation for deep multi-agent reinforcement learning,

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T05:22:10.184154Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T05:22:09.586181Z digest=sha256:e79306dae3e404b8318985495274d89570ab20e8caff771b8fa44fc78dacd78c

Observation 2d9f5118-7c3c-44df-8d55-2387f9483969 · outbound

This paper cites Counterfactual multi-agent policy gradients,.

Towards Fault Tolerance in Multi-Agent Reinforcement Learning Counterfactual multi-agent policy gradients,

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-12T05:22:09.590610Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T05:22:09.590610Z digest=sha256:687199f3de9ff099f39b2e31628292b2b6624b764de7c1695528209fc8aaa645

Observation 4b937b0c-95be-456f-a81c-23adbaeac636 · outbound

This paper cites Evolutionary population curriculum for scaling multi-agent reinforcement learning,.

Towards Fault Tolerance in Multi-Agent Reinforcement Learning Evolutionary population curriculum for scaling multi-agent reinforcement learning,

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T05:22:10.159221Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T05:22:09.595066Z digest=sha256:efcc30fef26a12af7683692ebae963f04c49998559b4411fe4f907d00fef01da

Observation 8c3d8768-4139-49f6-81e1-8c999dccd855 · outbound

This paper cites Scalable autonomous separation assurance with heterogeneous multi-agent reinforcement learning,.

Towards Fault Tolerance in Multi-Agent Reinforcement Learning Scalable autonomous separation assurance with heterogeneous multi-agent reinforcement learning,

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T05:22:10.144161Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T05:22:09.599722Z digest=sha256:674f3c9bc78bee0d86dcddaf07a9bcd973a2bae6b2b38d7705614e5d7a2fab64

Observation 6968ca23-d612-4128-9164-b8c59ebb838c · outbound

This paper cites Asynchronous multi-agent reinforcement learning for efficient real-time multi-robot cooperative exploration,.

Towards Fault Tolerance in Multi-Agent Reinforcement Learning Asynchronous multi-agent reinforcement learning for efficient real-time multi-robot cooperative exploration,

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T05:22:10.127665Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T05:22:09.604174Z digest=sha256:dc469aa7e2ab7a3712409f9912bc2312105d89ab816af67bf9ab191be31c31d4

Observation 70b76b4f-0de7-49e7-b5b3-02d52f148a6d · outbound

This paper cites Race: improve multi-agent reinforcement learning with representation asymmetry and collaborative evolution,.

Towards Fault Tolerance in Multi-Agent Reinforcement Learning Race: improve multi-agent reinforcement learning with representation asymmetry and collaborative evolution,

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T05:22:10.110615Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T05:22:09.608839Z digest=sha256:a4daf7b185a9258a45e1630cfa96950d1633de3bcba5c4194c9ed6a055daa753

Observation 495f4bb4-5859-4281-9b9d-fafc1d7c8091 · outbound

This paper cites Effective multi-agent deep reinforcement learning control with relative entropy regularization,.

Towards Fault Tolerance in Multi-Agent Reinforcement Learning Effective multi-agent deep reinforcement learning control with relative entropy regularization,

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T05:22:10.095039Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T05:22:09.613216Z digest=sha256:205fcb813846f4026ccd15aad66f68797b688077a5250eda596e93e163e70a8e

Observation c1eaa6fb-d84c-45e4-8f53-fcc45f27b21d · outbound

This paper cites Multi-task multi- agent reinforcement learning with task-entity transformers and value decomposition training,.

Towards Fault Tolerance in Multi-Agent Reinforcement Learning Multi-task multi- agent reinforcement learning with task-entity transformers and value decomposition training,

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T05:22:10.079621Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T05:22:09.618034Z digest=sha256:cc70ddb362921a5393fe4f523cdbeba2e1f2c7ec13889aa5f172c314b537ee77

Observation ea7b6dc6-ded2-43df-9f8c-b711af13fd8e · outbound

This paper cites R-MADDPG for Partially Observable Environments and Limited Communication.

Towards Fault Tolerance in Multi-Agent Reinforcement Learning R-MADDPG for Partially Observable Environments and Limited Communication

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-12T05:22:09.623794Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T05:22:09.623794Z digest=sha256:b6ccfcadb0759d3fe89d89c45ca8f30890975d0955bac41311495ce3bb6de82d

Observation bc86a9c6-c263-4a78-9d7f-06f5cf8d9b8c · outbound

This paper cites Multi-agent Deep Reinforcement Learning with Extremely Noisy Observations.

Towards Fault Tolerance in Multi-Agent Reinforcement Learning Multi-agent Deep Reinforcement Learning with Extremely Noisy Observations

Reference 35

Resolution
verified exact
local_arxiv, observed 2026-08-12T05:22:09.827043Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T05:22:09.628959Z digest=sha256:b4e92554b14372583246ae9f5657519792c881469bbe7e51e8fb84f00b74615e

Observation ee732337-c8b8-41d3-8e83-302355453081 · outbound

This paper cites A general survey on attention mechanisms in deep learning,.

Towards Fault Tolerance in Multi-Agent Reinforcement Learning A general survey on attention mechanisms in deep learning,

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T05:22:10.064923Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T05:22:09.634357Z digest=sha256:3da9a448ff71f9494570aaed7504f55263cb09f3d14b740e752947da60d6c9bd

Observation b1882652-b2bd-4433-ab91-72120af3ce3a · outbound

This paper cites Neural machine translation by jointly learning to align and translate,.

Towards Fault Tolerance in Multi-Agent Reinforcement Learning Neural machine translation by jointly learning to align and translate,

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T05:22:10.049561Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T05:22:09.639520Z digest=sha256:3a5522b8e9e6418d436440e15de960a036508923edc60e2dead206a7d90839b3

Observation 65d3c800-807b-4493-ade8-9bf6cab96d02 · outbound

This paper cites Attention is all you need,.

Towards Fault Tolerance in Multi-Agent Reinforcement Learning Attention is all you need,

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-12T05:22:09.644089Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T05:22:09.644089Z digest=sha256:fbbe258cbdecbdba02490d2a9654e5f45810a8674bd01f50b392bb03e22fe566

Observation 6ba2b20f-7798-412e-a1e8-2f7dad889f7e · outbound

This paper cites Large Language Models: A Survey.

Towards Fault Tolerance in Multi-Agent Reinforcement Learning Large Language Models: A Survey

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-12T05:22:09.649369Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T05:22:09.649369Z digest=sha256:4dcad0c6191d5e86aa70645cf7e6a0bc9176025a05a44e4e8b73152ff4f19144

Observation b2f587ee-4b3e-4c3a-bf79-827809965a3a · outbound

This paper cites Recurrent models of visual attention,.

Towards Fault Tolerance in Multi-Agent Reinforcement Learning Recurrent models of visual attention,

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T05:22:10.024852Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T05:22:09.654574Z digest=sha256:aaf1236f0e3ad1212250387b1c90d45a5b12cebda1deb3bcd8c5a696dfef5ef1

Observation c1d66f45-273c-4f6a-98d1-19d5fabc7689 · outbound

This paper cites An image is worth 16x16 words: Transformers for image recognition at scale,.

Towards Fault Tolerance in Multi-Agent Reinforcement Learning An image is worth 16x16 words: Transformers for image recognition at scale,

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-12T05:22:09.659264Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T05:22:09.659264Z digest=sha256:920624904741296988ea482fb06de7618abf35b1a68a80d42161f993fa5db215

Observation 191801cd-6a7f-4a8e-9feb-0ae4f20a7948 · outbound

This paper cites Actor-attention-critic for multi-agent reinforcement learning,.

Towards Fault Tolerance in Multi-Agent Reinforcement Learning Actor-attention-critic for multi-agent reinforcement learning,

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T05:22:09.997264Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T05:22:09.663893Z digest=sha256:1a1b0a08c1e4ea89834f674fea91a8cdbd1bd37d392197676a6a3ffad4add75d

Observation 1f890d80-f54e-4de1-adc4-a83312ed7550 · outbound

This paper cites Attention-based recurrence for multi-agent rein- forcement learning under stochastic partial observability,.

Towards Fault Tolerance in Multi-Agent Reinforcement Learning Attention-based recurrence for multi-agent rein- forcement learning under stochastic partial observability,

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T05:22:09.982142Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T05:22:09.668955Z digest=sha256:94580482451b2373cc95aa3b9e59c1aed2b599ad3ef61aad992ca1646ed761c0

Observation 80f40410-d8c9-4aaa-8315-3f267c55bccf · outbound

This paper cites Attention-Guided Contrastive Role Representations for Multi-Agent Reinforcement Learning.

Towards Fault Tolerance in Multi-Agent Reinforcement Learning Attention-Guided Contrastive Role Representations for Multi-Agent Reinforcement Learning

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-12T05:22:09.673402Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T05:22:09.673402Z digest=sha256:f085614f8c26b44ccf3a3132b7caf6f3e282589340a6d81283d820021afbeaab

Observation c8791e80-f72a-4d9d-bf5b-ae067ef72627 · outbound

This paper cites Novel distributed grus based on hybrid self-attention mechanism for dynamic soft sensing,.

Towards Fault Tolerance in Multi-Agent Reinforcement Learning Novel distributed grus based on hybrid self-attention mechanism for dynamic soft sensing,

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T05:22:09.967218Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T05:22:09.678574Z digest=sha256:de1587f84a546f65f7dc8d912effcc666f0a94e6088e4b965d937941db45e576

Observation 053d63a2-e868-499f-bca3-4a2781eded09 · outbound

This paper cites Learning from noisy labels with distillation,.

Towards Fault Tolerance in Multi-Agent Reinforcement Learning Learning from noisy labels with distillation,

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T05:22:09.952556Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T05:22:09.683251Z digest=sha256:250e6ff7cc7ee58116688f39d639f54260fae59c8f6ff40880934dcb6c2dbfef

Observation e55e301d-8c95-405f-92ef-f574644d8939 · outbound

This paper cites Curriculum reinforcement learning from avoiding collisions to navigating among movable obstacles in diverse environments,.

Towards Fault Tolerance in Multi-Agent Reinforcement Learning Curriculum reinforcement learning from avoiding collisions to navigating among movable obstacles in diverse environments,

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T05:22:09.937232Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T05:22:09.687886Z digest=sha256:0562e22d522e27788f6f0f1abb7b23d6fa0eb04167407980600640390d88b8f5

Observation 6abb803a-19f4-4760-9108-5abe83fcfddf · outbound

This paper cites Training region-based object detectors with online hard example mining,.

Towards Fault Tolerance in Multi-Agent Reinforcement Learning Training region-based object detectors with online hard example mining,

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-12T05:22:09.692721Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T05:22:09.692721Z digest=sha256:b3339b17d3079936a052f882aa0ac82b30e4af8daaa968bb6a06739ac5e719c6

Observation 25d69e2b-8179-4899-8af9-272d865f768b · outbound

This paper cites Markov games as a framework for multi-agent rein- forcement learning,.

Towards Fault Tolerance in Multi-Agent Reinforcement Learning Markov games as a framework for multi-agent rein- forcement learning,

Reference 49

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T05:22:09.913197Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T05:22:09.697574Z digest=sha256:6dd0cdef1b48e4fde50654f6a59da22fdef8e2c05d7beec33bb29d86a0df7e34

Observation 09988ba2-3786-41d3-b537-8435bf29f365 · outbound

This paper cites Planning and acting in partially observable stochastic domains,.

Towards Fault Tolerance in Multi-Agent Reinforcement Learning Planning and acting in partially observable stochastic domains,

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T05:22:09.898186Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T05:22:09.702321Z digest=sha256:7a90be0e1c93dc054470a3f6997f5953402164be1b20d46458d08f6047bfc275

Observation 26dde8a6-cab0-4f40-ac81-0afd5993a78b · outbound

This paper cites Continuous control with deep reinforcement learning,.

Towards Fault Tolerance in Multi-Agent Reinforcement Learning Continuous control with deep reinforcement learning,

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T05:22:09.883105Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T05:22:09.706728Z digest=sha256:00319905b56204e6c408fb43deb172068bd58e85d4c5cfc6af24db740e4d749b

Observation 6818dc0f-6a56-4ce8-8f52-27eb888893bb · outbound

This paper cites RLlib: Abstractions for Distributed Reinforcement Learning.

Towards Fault Tolerance in Multi-Agent Reinforcement Learning RLlib: Abstractions for Distributed Reinforcement Learning

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-12T05:22:09.711386Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T05:22:09.711386Z digest=sha256:ca723d8c28ae2a7a153573ade12db32e3eb9e49a3337683834fc1d1b9191e4b2

Observation 5130b66b-b7bf-41fc-a2e8-19851e9801e3 · outbound

This paper cites rlpyt: A Research Code Base for Deep Reinforcement Learning in PyTorch.

Towards Fault Tolerance in Multi-Agent Reinforcement Learning rlpyt: A Research Code Base for Deep Reinforcement Learning in PyTorch

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-12T05:22:09.716785Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T05:22:09.716785Z digest=sha256:3172665d7e4435d0ccff916f97629980a1df8dcad8a0c3c96827d067155728b2

Observation 0769e7bb-b4a0-4c92-8473-4d708b32184f · outbound

This paper cites Experiment tracking with weights and biases,.

Towards Fault Tolerance in Multi-Agent Reinforcement Learning Experiment tracking with weights and biases,

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-12T05:22:09.722683Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T05:22:09.722683Z digest=sha256:7bc70aa57eaf8704cd318a37f9da13fc8023c0eef02c97c7177b1d7a3b94c37b

Pith citing papers

Observation ac8229a5-852d-41d9-ad84-737d7149333d · inbound

Exploring Critical Testing Scenarios for Decision-Making Policies: An LLM Approach cites this paper.

Exploring Critical Testing Scenarios for Decision-Making Policies: An LLM Approach Towards Fault Tolerance in Multi-Agent Reinforcement Learning

Reference 10

Resolution
verified exact
local_arxiv, observed 2026-08-11T19:28:52.644424Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-11T19:28:52.274750Z digest=sha256:5eb58d0b0d91faad78c030ebdb5fdb4e4de74588f0beb7989f40db95ea603ae2