Pith. sign in

Paper Citation Record · LEDGER

Fast Multi-Agent Temporal-Difference Learning via Homotopy Stochastic Primal-Dual Optimization

As of 22 August 2026, this Paper Citation Record lists 58 of 58 outbound references and 0 inbound Pith citation observations for arXiv:1908.02805.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
1908.02805 v4

Coverage vector

measured 58 of 58 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-14T14:42:28.270299Z

measured 58 of 58 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-22T06:32:14.747728+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

58 of 58 outbound references displayed

  • verified exact6
  • verified fuzzy46
  • unresolved6
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation e15827cd-0928-4329-b202-777b519b224e · outbound

This paper cites an unresolved cited work.

Fast Multi-Agent Temporal-Difference Learning via Homotopy Stochastic Primal-Dual Optimization Unresolved cited work

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-14T14:42:28.021233Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T14:42:28.021233Z digest=sha256:07abfac9fbd9a5d10ea1481366328edca16dc81931252c64af3d945b1fe0f0ec

Observation 28d695ab-85b1-4a7f-9e67-87ff82746cfb · outbound

This paper cites Learning to predict by the methods of temporal differences,.

Fast Multi-Agent Temporal-Difference Learning via Homotopy Stochastic Primal-Dual Optimization Learning to predict by the methods of temporal differences,

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-14T14:42:28.026903Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T14:42:28.026903Z digest=sha256:e0dba64b8c21c30097ae00f8a6314e1ba6d70fff6ccab54b6e3b90b5f0b95e5c

Observation b85d0643-9bca-4e5c-aaf1-4cba95ea484a · outbound

This paper cites an unresolved cited work.

Fast Multi-Agent Temporal-Difference Learning via Homotopy Stochastic Primal-Dual Optimization Unresolved cited work

Reference 3

Resolution
unresolved
raw_fallback, observed 2026-08-14T14:42:29.156338Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-14T14:42:28.031474Z digest=sha256:597364fb2fdfddbb6abc2339eb4560b881f571b9abbf46b830701ff85bdb2e08

Observation f097c121-60d1-48e4-9657-2d4f9196d5ed · outbound

This paper cites Residual algorithms: reinforcement learning with function approximation,.

Fast Multi-Agent Temporal-Difference Learning via Homotopy Stochastic Primal-Dual Optimization Residual algorithms: reinforcement learning with function approximation,

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T14:42:29.142212Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-14T14:42:28.035942Z digest=sha256:27db9769420b5f5e30b4062a38f64b988b9a75bf6a26683bafcd39fe34862142

Observation 012a80d8-3e06-4301-8e8d-96b9f264a07b · outbound

This paper cites Human-level control through deep reinforcement learning,.

Fast Multi-Agent Temporal-Difference Learning via Homotopy Stochastic Primal-Dual Optimization Human-level control through deep reinforcement learning,

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T14:42:29.128048Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-14T14:42:28.040138Z digest=sha256:93ea7b3e1c63a8a6ef660208fa463cc88fa8f243767ace4fb9c64e52abd823f9

Observation 7fa4b313-3b5a-4079-8380-02ec11b83324 · outbound

This paper cites Mastering the game of Go with deep neural networks and tree search,.

Fast Multi-Agent Temporal-Difference Learning via Homotopy Stochastic Primal-Dual Optimization Mastering the game of Go with deep neural networks and tree search,

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T14:42:29.113770Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-14T14:42:28.044430Z digest=sha256:49f249cc9ca6192828fea46be25dac62e5e715eadd6028210b2e0177603efe39

Observation 92d3b55c-5b72-454e-ac5f-94f028458fe6 · outbound

This paper cites Residential energy management in smart grid: A Markov decision process-based approach,.

Fast Multi-Agent Temporal-Difference Learning via Homotopy Stochastic Primal-Dual Optimization Residential energy management in smart grid: A Markov decision process-based approach,

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T14:42:29.098963Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-14T14:42:28.048757Z digest=sha256:7bf0fe1d7eb34aabfee9c38eb68962a338561983397b65c3734650367a86410a

Observation e95ae937-a098-4317-aebf-beff5d4a8211 · outbound

This paper cites Multiagent reinforcement learning for urban traffic control using coordination graphs,.

Fast Multi-Agent Temporal-Difference Learning via Homotopy Stochastic Primal-Dual Optimization Multiagent reinforcement learning for urban traffic control using coordination graphs,

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T14:42:29.083452Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-14T14:42:28.052409Z digest=sha256:60ff22e1ed54243c42800089db4cafffad9676a80994dc2ffb485617570b2480

Observation fbb5bb53-4125-4ee2-b78b-e8f348687ee4 · outbound

This paper cites A distributed actor-critic algorithm and applications to mobile sensor network coordination problems,.

Fast Multi-Agent Temporal-Difference Learning via Homotopy Stochastic Primal-Dual Optimization A distributed actor-critic algorithm and applications to mobile sensor network coordination problems,

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T14:42:29.067997Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-14T14:42:28.056249Z digest=sha256:b13f8a4cf24782704edcaf2d6f105ede81311b29ec1aab0e7bc10a136781e25d

Observation 49766629-fc3d-454e-9cc1-9b5052ce5a04 · outbound

This paper cites Reinforcement learning in robotics: A survey,.

Fast Multi-Agent Temporal-Difference Learning via Homotopy Stochastic Primal-Dual Optimization Reinforcement learning in robotics: A survey,

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T14:42:29.053696Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-14T14:42:28.060569Z digest=sha256:5221724caeeb3d6cdeacafddf2f03e39800da3112ff6ff84f168410d760c1eb6

Observation 9b700b1d-9cf8-455e-b181-e9d553b1c303 · outbound

This paper cites Distributed policy evaluation under multiple behavior strategies,.

Fast Multi-Agent Temporal-Difference Learning via Homotopy Stochastic Primal-Dual Optimization Distributed policy evaluation under multiple behavior strategies,

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T14:42:29.040159Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-14T14:42:28.065089Z digest=sha256:ec4470b07ed3f1296af33bba069a7fb89f596bd80a6f1c99786967ee127b77ab

Observation b2b6d37e-4d05-4f0a-bb5e-c999ed00be54 · outbound

This paper cites Distributed reinforcement learning via gossip,.

Fast Multi-Agent Temporal-Difference Learning via Homotopy Stochastic Primal-Dual Optimization Distributed reinforcement learning via gossip,

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T14:42:29.026525Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-14T14:42:28.069673Z digest=sha256:a2d9af3dfd42b1f0b971a84a207f69927eb15ccef3a40431d1a65015c7157aa5

Observation a3990ef9-fa7d-433a-a8e8-0f316331d07f · outbound

This paper cites Primal-dual algorithm for distributed reinforcement learning: distributed GTD,.

Fast Multi-Agent Temporal-Difference Learning via Homotopy Stochastic Primal-Dual Optimization Primal-dual algorithm for distributed reinforcement learning: distributed GTD,

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T14:42:29.012073Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-14T14:42:28.073776Z digest=sha256:bd277b3b69f07d91f8f37b0021a4e25ca01531cf509c8d1e3791a1558421701e

Observation 517665fd-0af4-4b0c-9e51-34bec6416c86 · outbound

This paper cites Multi-agent reinforcement learning via double averaging primal-dual optimization,.

Fast Multi-Agent Temporal-Difference Learning via Homotopy Stochastic Primal-Dual Optimization Multi-agent reinforcement learning via double averaging primal-dual optimization,

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T14:42:28.997808Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-14T14:42:28.078023Z digest=sha256:30ad04b1aa3cf626fcffd8f6545ecf282be87cb709933cc146725b475b541e82

Observation f08ca55e-3645-4e75-8322-1e38b754e677 · outbound

This paper cites Multi-Agent Fully Decentralized Value Function Learning with Linear Convergence Rates.

Fast Multi-Agent Temporal-Difference Learning via Homotopy Stochastic Primal-Dual Optimization Multi-Agent Fully Decentralized Value Function Learning with Linear Convergence Rates

Reference 15

Resolution
verified exact
local_arxiv, observed 2026-08-14T14:42:28.457039Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-14T14:42:28.082359Z digest=sha256:fa01c11d45212a11635befc5e3524aec7e815b7c48d2c74b83ae6a3485d1c68e

Observation 5a347f68-9944-4d32-b06b-89d9b2ca8741 · outbound

This paper cites Finite-time analysis of distributed TD(0) with linear function approximation on multi-agent reinforcement learning,.

Fast Multi-Agent Temporal-Difference Learning via Homotopy Stochastic Primal-Dual Optimization Finite-time analysis of distributed TD(0) with linear function approximation on multi-agent reinforcement learning,

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T14:42:28.983791Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-14T14:42:28.086840Z digest=sha256:4467b89620fc3c7d853e543e380aae116217082a67aecfe5ccc156f522352b4f

Observation d057ba14-535b-4eea-a1e0-a34e572e6da5 · outbound

This paper cites Finite-Time Performance of Distributed Temporal Difference Learning with Linear Function Approximation.

Fast Multi-Agent Temporal-Difference Learning via Homotopy Stochastic Primal-Dual Optimization Finite-Time Performance of Distributed Temporal Difference Learning with Linear Function Approximation

Reference 17

Resolution
verified exact
local_arxiv, observed 2026-08-14T14:42:28.436052Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-14T14:42:28.091067Z digest=sha256:5a8abf8c167e5e1b4f4573226ff8b9bb3b875c855ffbd5a9583acd848fee561f

Observation 2f048e6b-ac78-4cbd-8448-70eadbe156a0 · outbound

This paper cites Finite-Sample Analysis of Decentralized Temporal-Difference Learning with Linear Function Approximation.

Fast Multi-Agent Temporal-Difference Learning via Homotopy Stochastic Primal-Dual Optimization Finite-Sample Analysis of Decentralized Temporal-Difference Learning with Linear Function Approximation

Reference 18

Resolution
verified exact
local_arxiv, observed 2026-08-14T14:42:28.415808Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-14T14:42:28.095985Z digest=sha256:57820f78d25a4738a4b3c5b6b8fe3048dc88f7ee44fd0148bbdb6d323c1d2787

Observation 59102187-fb98-4eef-81bb-b01ed09342aa · outbound

This paper cites Fully Asynchronous Policy Evaluation in Distributed Reinforcement Learning over Networks.

Fast Multi-Agent Temporal-Difference Learning via Homotopy Stochastic Primal-Dual Optimization Fully Asynchronous Policy Evaluation in Distributed Reinforcement Learning over Networks

Reference 19

Resolution
verified exact
local_arxiv, observed 2026-08-14T14:42:28.392356Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-14T14:42:28.100700Z digest=sha256:47234c2283be60d345392c6f1472000e6e4a1c29b5f4902fb545c795f2b6d1a4

Observation f0da8c8e-5c83-45ba-a38c-8d6038b48645 · outbound

This paper cites An analysis of temporal-difference learning with function approximation,.

Fast Multi-Agent Temporal-Difference Learning via Homotopy Stochastic Primal-Dual Optimization An analysis of temporal-difference learning with function approximation,

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T14:42:28.969568Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-14T14:42:28.105575Z digest=sha256:cbb5e9e5f496fe6a3117d45b6d6d8817d4582d6e83b55690b47f832f54a5a46a

Observation 4e0211f5-e45b-4b76-a0f6-468639555184 · outbound

This paper cites Szepesv ´ari, Algorithms for Reinforcement Learning.

Fast Multi-Agent Temporal-Difference Learning via Homotopy Stochastic Primal-Dual Optimization Szepesv ´ari, Algorithms for Reinforcement Learning

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T14:42:28.954808Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-14T14:42:28.110004Z digest=sha256:e9b554b2af19666fa1401ba830d9d226a186f0302cb1e15a98a53021717f8cff

Observation 73fcb46f-c3b1-4e43-866c-c2630f4dd833 · outbound

This paper cites Linear least-squares algorithms for temporal difference learning,.

Fast Multi-Agent Temporal-Difference Learning via Homotopy Stochastic Primal-Dual Optimization Linear least-squares algorithms for temporal difference learning,

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T14:42:28.940962Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-14T14:42:28.114330Z digest=sha256:ccbc34d0ad2e85a14b6cef6508170d29885395045b498ccbc489fa8892be2611

Observation f90470b3-a611-4d59-91af-bc6f3ee59f2d · outbound

This paper cites A convergent O(n) temporal-difference algorithm for off-policy learning with linear function approximation,.

Fast Multi-Agent Temporal-Difference Learning via Homotopy Stochastic Primal-Dual Optimization A convergent O(n) temporal-difference algorithm for off-policy learning with linear function approximation,

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T14:42:28.927720Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-14T14:42:28.118634Z digest=sha256:20cfa397b4ae2592191d558cad5bfd4cb17a6387e7c45f252e2f12239fa4a33d

Observation ac13deb6-978c-4372-8152-6ae0dd191a98 · outbound

This paper cites Fast gradient-descent methods for temporal- difference learning with linear function approximation,.

Fast Multi-Agent Temporal-Difference Learning via Homotopy Stochastic Primal-Dual Optimization Fast gradient-descent methods for temporal- difference learning with linear function approximation,

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T14:42:28.914294Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-14T14:42:28.122954Z digest=sha256:477dfbf59086fda8b82cca41865aaa081a8c42e771bc8c38d24046e01684f3c1

Observation 6f565f6f-6aa6-4de7-9fa3-e22375eeb500 · outbound

This paper cites The ode method for convergence of stochastic approximation and reinforcement learning,.

Fast Multi-Agent Temporal-Difference Learning via Homotopy Stochastic Primal-Dual Optimization The ode method for convergence of stochastic approximation and reinforcement learning,

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T14:42:28.900843Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-14T14:42:28.127491Z digest=sha256:5933691f3027bd51cf63b1ba123a45718edff0e4197a2add022b0efd75bb2ebe

Observation 20c35761-fc08-47d0-90af-40a44c7d1cc0 · outbound

This paper cites Finite sample analyses for TD (0) with function approximation,.

Fast Multi-Agent Temporal-Difference Learning via Homotopy Stochastic Primal-Dual Optimization Finite sample analyses for TD (0) with function approximation,

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T14:42:28.884765Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-14T14:42:28.132222Z digest=sha256:db739139c6e9d72ffd07f8f7b60b963a3d05ca8b3625a6f8ff185dff1a379f10

Observation 6939bf3f-a554-4063-a6a5-98a4180b16b3 · outbound

This paper cites Finite sample analysis of two-time scale stochastic approximation with applications to reinforcement learning,.

Fast Multi-Agent Temporal-Difference Learning via Homotopy Stochastic Primal-Dual Optimization Finite sample analysis of two-time scale stochastic approximation with applications to reinforcement learning,

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T14:42:28.868871Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-14T14:42:28.136958Z digest=sha256:fedf0578557593d1f05cebfbe1f3c40f0df7d04c66400f4f487af067f28abf0a

Observation 66367c32-08dd-40ce-8b9a-ab70528672e5 · outbound

This paper cites Linear stochastic approximation: How far does constant step-size and iterate averaging go?.

Fast Multi-Agent Temporal-Difference Learning via Homotopy Stochastic Primal-Dual Optimization Linear stochastic approximation: How far does constant step-size and iterate averaging go?

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T14:42:28.853601Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-14T14:42:28.141580Z digest=sha256:640b68be70ac752d23bd0029a70136b976ff1c5fbc2b049dbe86f28b656534c1

Observation 06eb515a-d082-401a-80c6-2b3e17f9add6 · outbound

This paper cites A finite time analysis of temporal difference learning with linear function approximation,.

Fast Multi-Agent Temporal-Difference Learning via Homotopy Stochastic Primal-Dual Optimization A finite time analysis of temporal difference learning with linear function approximation,

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T14:42:28.839187Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-14T14:42:28.145862Z digest=sha256:2ed0e706f92c634a2ed5c19d28516f382088367dc9162f35c20c0dc4537556fc

Observation 95e39187-8e19-4362-8185-d6a6c666981e · outbound

This paper cites Finite-time error bounds for linear stochastic approximation and TD learning,.

Fast Multi-Agent Temporal-Difference Learning via Homotopy Stochastic Primal-Dual Optimization Finite-time error bounds for linear stochastic approximation and TD learning,

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T14:42:28.825063Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-14T14:42:28.150011Z digest=sha256:3be6e17e592f1d6e889d2f5aecc3a11c842c113805de87ddf924bca41706c028

Observation 45ba3c6c-be93-46d6-b446-428f7ee9b6b2 · outbound

This paper cites Finite-sample analysis for SARSA and Q-learning with linear function approximation,.

Fast Multi-Agent Temporal-Difference Learning via Homotopy Stochastic Primal-Dual Optimization Finite-sample analysis for SARSA and Q-learning with linear function approximation,

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T14:42:28.810675Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-14T14:42:28.154518Z digest=sha256:f1459267ad6c63d472426506ccb66cc6b03e40b9a73ca58880b1b24fd8aba829

Observation 51467f7d-61b6-49b9-8f4f-a80679da323c · outbound

This paper cites Characterizing the exact behaviors of temporal difference learning algorithms using markov jump linear system theory,.

Fast Multi-Agent Temporal-Difference Learning via Homotopy Stochastic Primal-Dual Optimization Characterizing the exact behaviors of temporal difference learning algorithms using markov jump linear system theory,

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T14:42:28.794956Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-14T14:42:28.158742Z digest=sha256:a9c405e7808d83228c440d23289900df2c8a0a3d66d28eab31f0d3c26b405d5f

Observation ef182439-bc13-4652-9d4d-4822cfd586f7 · outbound

This paper cites Finite-sample analysis of proximal gradient TD algorithms,.

Fast Multi-Agent Temporal-Difference Learning via Homotopy Stochastic Primal-Dual Optimization Finite-sample analysis of proximal gradient TD algorithms,

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T14:42:28.779789Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-14T14:42:28.163018Z digest=sha256:a6847e1d742d9917e82632e8d84a629bb0ff255311cbff164b6ac3044b05110f

Observation b780a0a6-e9a2-45ac-b7ea-a3f85231aff6 · outbound

This paper cites Robust stochastic approximation approach to stochastic programming,.

Fast Multi-Agent Temporal-Difference Learning via Homotopy Stochastic Primal-Dual Optimization Robust stochastic approximation approach to stochastic programming,

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T14:42:28.764325Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-14T14:42:28.166628Z digest=sha256:819a0dacd1dfc807367b4121bb38f7cefdb6224ffb8bdffb31957f704410fa33

Observation a4617d73-7f75-4c8e-b1a1-63a25af517d5 · outbound

This paper cites Finite sample analysis of the GTD policy evaluation algorithms in Markov setting,.

Fast Multi-Agent Temporal-Difference Learning via Homotopy Stochastic Primal-Dual Optimization Finite sample analysis of the GTD policy evaluation algorithms in Markov setting,

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T14:42:28.751541Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-14T14:42:28.170230Z digest=sha256:b3e12d5eda1a8b65448bb061f7470776ecf425247b68866142cdc51d45263ae0

Observation ac25571a-6ed9-426b-8388-51be3f3b9244 · outbound

This paper cites Convergent TREE BACKUP and RETRACE with function approximation,.

Fast Multi-Agent Temporal-Difference Learning via Homotopy Stochastic Primal-Dual Optimization Convergent TREE BACKUP and RETRACE with function approximation,

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T14:42:28.738015Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-14T14:42:28.174649Z digest=sha256:2265f72972ad8cab257c7d7d7467d897ac89c9596787406ee68e24960ec99add

Observation 73ea2fba-c87f-4977-903a-97b03dd3cf02 · outbound

This paper cites Multi-agent temporal-difference learning with linear function approximation: Weak convergence under time-varying network topologies,.

Fast Multi-Agent Temporal-Difference Learning via Homotopy Stochastic Primal-Dual Optimization Multi-agent temporal-difference learning with linear function approximation: Weak convergence under time-varying network topologies,

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T14:42:28.723701Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-14T14:42:28.178323Z digest=sha256:14ce1a0e769288c6e3f29b1bf9459ff64a25a02237f8398b3ab0e2d35603998f

Observation e08e85de-756e-4c96-a584-477976830884 · outbound

This paper cites Gossip algorithms: Design, analysis and applications,.

Fast Multi-Agent Temporal-Difference Learning via Homotopy Stochastic Primal-Dual Optimization Gossip algorithms: Design, analysis and applications,

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T14:42:28.708438Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-14T14:42:28.182268Z digest=sha256:fe74460ec74a1f50ff97e76247768ee1736846bbd4e86b3d3f66a0cb9cbd031f

Observation 1f620643-cab7-4efa-aa6c-0eb9ee15de60 · outbound

This paper cites Finite-Time Performance of Distributed Two-Time-Scale Stochastic Approximation.

Fast Multi-Agent Temporal-Difference Learning via Homotopy Stochastic Primal-Dual Optimization Finite-Time Performance of Distributed Two-Time-Scale Stochastic Approximation

Reference 39

Resolution
verified exact
local_arxiv, observed 2026-08-14T14:42:28.371015Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-14T14:42:28.187587Z digest=sha256:b283a44d7ade2ae76444122c4c79b716c2ff9d1487a9e5f672fddad8aa5f4fa3

Observation b2d4b097-a69b-4af9-8cd6-1134b27a7fa0 · outbound

This paper cites On the averaged stochastic approximation for linear regression,.

Fast Multi-Agent Temporal-Difference Learning via Homotopy Stochastic Primal-Dual Optimization On the averaged stochastic approximation for linear regression,

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T14:42:28.693809Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-14T14:42:28.191638Z digest=sha256:f7f063124fe4329f84dae65be4fb1eeea0af0bb4781cbe912650499c35f600a8

Observation e3bf95e0-e298-4e82-b8c1-75be7ed12b99 · outbound

This paper cites Fully decentralized multi-agent reinforcement learning with networked agents,.

Fast Multi-Agent Temporal-Difference Learning via Homotopy Stochastic Primal-Dual Optimization Fully decentralized multi-agent reinforcement learning with networked agents,

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T14:42:28.678415Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-14T14:42:28.195009Z digest=sha256:901170d8a41d4d7764199c33d11b244eee642745171fb7a5cdec9950f089c013

Observation 6eb6318d-1818-47b9-8b8a-758f95e86eb9 · outbound

This paper cites Multi-Agent Reinforcement Learning: A Selective Overview of Theories and Algorithms.

Fast Multi-Agent Temporal-Difference Learning via Homotopy Stochastic Primal-Dual Optimization Multi-Agent Reinforcement Learning: A Selective Overview of Theories and Algorithms

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-14T14:42:28.199222Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T14:42:28.199222Z digest=sha256:27cc3bbbffd337a0318ba4b101697f5f5a22a15d0f770f7d9bf7da0b6665dcaf

Observation 15053b9a-4992-476a-9986-4bf793619ce4 · outbound

This paper cites Decentralized Multi-Agent Reinforcement Learning with Networked Agents: Recent Advances.

Fast Multi-Agent Temporal-Difference Learning via Homotopy Stochastic Primal-Dual Optimization Decentralized Multi-Agent Reinforcement Learning with Networked Agents: Recent Advances

Reference 43

Resolution
verified exact
local_arxiv, observed 2026-08-14T14:42:28.332117Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-14T14:42:28.204080Z digest=sha256:a9ea554951a6a1c245464059d5171e8523451dfdc31f1da6bcf75cb58a8b8e8f

Observation 331c0120-082c-4297-934e-f0b439308d19 · outbound

This paper cites Optimization for reinforcement learning: From a single agent to cooperative agents,.

Fast Multi-Agent Temporal-Difference Learning via Homotopy Stochastic Primal-Dual Optimization Optimization for reinforcement learning: From a single agent to cooperative agents,

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T14:42:28.661986Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-14T14:42:28.208814Z digest=sha256:42ece9b9dd7ebb5aff0d253b40eb0c777425dc512dc555709981cc8c325796f1

Observation 96093c6c-b5ea-4845-a011-2fd34165f988 · outbound

This paper cites Dual averaging for distributed optimization: Convergence analysis and network scaling,.

Fast Multi-Agent Temporal-Difference Learning via Homotopy Stochastic Primal-Dual Optimization Dual averaging for distributed optimization: Convergence analysis and network scaling,

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T14:42:28.646492Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-14T14:42:28.212999Z digest=sha256:9c591e612d513c702580648c37b50d2fbbb090418db41c7a978a42fccfb554b0

Observation 322c32f7-041e-45b6-a7ab-24056a097dee · outbound

This paper cites A proximal-gradient homotopy method for the sparse least-squares problem,.

Fast Multi-Agent Temporal-Difference Learning via Homotopy Stochastic Primal-Dual Optimization A proximal-gradient homotopy method for the sparse least-squares problem,

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T14:42:28.629867Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-14T14:42:28.217396Z digest=sha256:c234e057825102d531b76e17e6033a67894be4404b8997435b9a56f2348681ae

Observation 2e4ff128-f5a6-4c47-98e5-4da431be0da5 · outbound

This paper cites Information-theoretic lower bounds on the oracle complexity of convex optimization,.

Fast Multi-Agent Temporal-Difference Learning via Homotopy Stochastic Primal-Dual Optimization Information-theoretic lower bounds on the oracle complexity of convex optimization,

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T14:42:28.615945Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-14T14:42:28.221610Z digest=sha256:00462e058221e5af23ed9e9d499104730e1ebaa11475faa8329ffb03d1d5ca40

Observation d71e7535-4f8d-4caf-a653-28f8b97eda96 · outbound

This paper cites Primal-Dual Distributed Temporal Difference Learning.

Fast Multi-Agent Temporal-Difference Learning via Homotopy Stochastic Primal-Dual Optimization Primal-Dual Distributed Temporal Difference Learning

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-14T14:42:28.226368Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T14:42:28.226368Z digest=sha256:acfa2497ba559781734958df9fb5a4c6ea88e50b3fe18803f0cc35a868908de5

Observation d20b27ad-6586-460c-9ef3-4e68192c65e8 · outbound

This paper cites Ergodic mirror descent,.

Fast Multi-Agent Temporal-Difference Learning via Homotopy Stochastic Primal-Dual Optimization Ergodic mirror descent,

Reference 49

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T14:42:28.600588Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-14T14:42:28.230720Z digest=sha256:40c41d41b80948c54ee07c9c321127f0000fad73eed0fdbf11b7b66ab6398ac2

Observation 44f75023-312d-4267-a165-1604f6da2615 · outbound

This paper cites an unresolved cited work.

Fast Multi-Agent Temporal-Difference Learning via Homotopy Stochastic Primal-Dual Optimization Unresolved cited work

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-14T14:42:28.234977Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T14:42:28.234977Z digest=sha256:9cf918e28bde47c0373a2aec1fbc3e07ff7eee266d96fa53d16041531d93b314

Observation 08f1fd21-5507-4b8b-aed7-64c44c070a52 · outbound

This paper cites Distributed subgradient methods for multi-agent optimization,.

Fast Multi-Agent Temporal-Difference Learning via Homotopy Stochastic Primal-Dual Optimization Distributed subgradient methods for multi-agent optimization,

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T14:42:28.574614Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-14T14:42:28.239573Z digest=sha256:c8b2dd6b18b5eb1f052a9a3ca50e05d0f7a0708dd83305864dac9557b1d47ca3

Observation e3dc3f03-1020-4ec2-b5d0-00c6b926807e · outbound

This paper cites Distributed strongly convex optimization,.

Fast Multi-Agent Temporal-Difference Learning via Homotopy Stochastic Primal-Dual Optimization Distributed strongly convex optimization,

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T14:42:28.560560Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-14T14:42:28.244768Z digest=sha256:9530afb2faa34b8db4efd4f363a5e6ddebfca455c5effc9c3df5af37e0aea9d9

Observation 0ce185fb-4ea1-4eb2-9816-287ff1cb412e · outbound

This paper cites RSG: Beating subgradient method without smoothness and strong convexity,.

Fast Multi-Agent Temporal-Difference Learning via Homotopy Stochastic Primal-Dual Optimization RSG: Beating subgradient method without smoothness and strong convexity,

Reference 53

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T14:42:28.545976Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-14T14:42:28.249153Z digest=sha256:9978a33c91afb00caba8bde95aee505bdf6a194dd28eba2bd447b493e5dc377a

Observation e032b661-ff55-49d1-a624-f209749af128 · outbound

This paper cites Homotopy smoothing for non-smooth problems with lower complexity than O (1/ϵ),.

Fast Multi-Agent Temporal-Difference Learning via Homotopy Stochastic Primal-Dual Optimization Homotopy smoothing for non-smooth problems with lower complexity than O (1/ϵ),

Reference 54

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T14:42:28.530768Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-14T14:42:28.253551Z digest=sha256:efeb2ec99aa70eaf30de2c4d1aac42891d6d3578a76884c837034255cc6397c3

Observation 0ed1903b-d18b-4bdb-a93c-06ad7f2944f6 · outbound

This paper cites Solving non-smooth constrained programs with lower complexity than O (1/ε): A primal-dual homotopy smoothing approach,.

Fast Multi-Agent Temporal-Difference Learning via Homotopy Stochastic Primal-Dual Optimization Solving non-smooth constrained programs with lower complexity than O (1/ε): A primal-dual homotopy smoothing approach,

Reference 55

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T14:42:28.515981Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-14T14:42:28.257995Z digest=sha256:0150c7966760bf8d3595ddaa27e5aac6a5da360ed69ddea02091b5c080603b81

Observation f47f105f-00c4-48e8-8a6e-76ebdf94aca7 · outbound

This paper cites Online convex programming and generalized infinitesimal gradient ascent,.

Fast Multi-Agent Temporal-Difference Learning via Homotopy Stochastic Primal-Dual Optimization Online convex programming and generalized infinitesimal gradient ascent,

Reference 56

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T14:42:28.500723Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-14T14:42:28.261827Z digest=sha256:0f0b7c29f9b401eb9097cbbc4e225c6e82e9954a4e02708d806faf8b5ff2df39

Observation 20700199-4016-4498-9da4-b56091d6d5f1 · outbound

This paper cites Subgradient methods for saddle-point problems,.

Fast Multi-Agent Temporal-Difference Learning via Homotopy Stochastic Primal-Dual Optimization Subgradient methods for saddle-point problems,

Reference 57

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T14:42:28.486704Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-14T14:42:28.265746Z digest=sha256:2cc748f456573182bb01d91647e33f4e131c5f8397560d82ea7aa28e6ece55cf

Observation 66a736dc-51d0-4107-95b9-88c0bfe1fc8f · outbound

This paper cites Optimum bounds for the distributions of martingales in Banach spaces,.

Fast Multi-Agent Temporal-Difference Learning via Homotopy Stochastic Primal-Dual Optimization Optimum bounds for the distributions of martingales in Banach spaces,

Reference 58

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T14:42:28.472142Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-14T14:42:28.270299Z digest=sha256:5e4f65fd275b3eff84cad73d7e26bb23eed8d260c865a742e13756f31c0c26d2

Pith citing papers

No inbound Pith citation observations are available.