Pith. sign in

Paper Citation Record · LEDGER

Variance-Reduced Q-Learning over Static and Time-Varying Networks

As of 7 August 2026, this Paper Citation Record lists 21 of 21 outbound references and 0 inbound Pith citation observations for arXiv:2607.21876.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2607.21876 v1

Coverage vector

measured 21 of 21 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-01T06:37:06.562214Z

measured 21 of 21 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

21 of 21 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved21
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 364cb31d-e2f0-46e0-9f2a-5d6a4c24dd1b · outbound

This paper cites Q-learning,.

Variance-Reduced Q-Learning over Static and Time-Varying Networks Q-learning,

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-01T06:37:06.223565Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T06:37:06.223565Z digest=sha256:c435d9f42131cc1962a811c1d55337cf13f19be81dde35633bde5b3a1eee9e53

Observation 2b7595d7-a1e4-4d5c-adec-72010cefbc01 · outbound

This paper cites QD-learning: A collaborative distributed strategy for multi-agent reinforcement learning through consensus + innovations,.

Variance-Reduced Q-Learning over Static and Time-Varying Networks QD-learning: A collaborative distributed strategy for multi-agent reinforcement learning through consensus + innovations,

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-01T06:37:06.319740Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T06:37:06.319740Z digest=sha256:940b830053208f8e5e6578f35046722627ef8451b6291c95f6352c03b00b2f36

Observation f2a38a6f-91a1-4ee0-844a-52b11e5d327d · outbound

This paper cites Fully decen- tralized multi-agent reinforcement learning with networked agents,.

Variance-Reduced Q-Learning over Static and Time-Varying Networks Fully decen- tralized multi-agent reinforcement learning with networked agents,

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-01T06:37:06.389659Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T06:37:06.389659Z digest=sha256:72e2e149698b9b02daf03f0430e15e3ef0d61d5ef35b567d4993a64056b41c17

Observation 8b4937c3-33de-4aa2-8131-2022f47eeba5 · outbound

This paper cites Distributed off-policy actor-critic reinforcement learning with policy consensus,.

Variance-Reduced Q-Learning over Static and Time-Varying Networks Distributed off-policy actor-critic reinforcement learning with policy consensus,

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-01T06:37:06.395768Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T06:37:06.395768Z digest=sha256:844adbac4d6dec25aebf0cbbc239f64e9bd2f3f6f6893ff694cb3a6f6ae43841

Observation 54045625-4abe-4239-8255-f0649625d93c · outbound

This paper cites Finite-time analysis of distributed TD (0) with linear function approximation on multi- agent reinforcement learning,.

Variance-Reduced Q-Learning over Static and Time-Varying Networks Finite-time analysis of distributed TD (0) with linear function approximation on multi- agent reinforcement learning,

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-01T06:37:06.404899Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T06:37:06.404899Z digest=sha256:1ee6d4d73e126cd23d2336841930431c58d2c19b2ab45f0e0c851734270e82b6

Observation 121c0001-aba4-494e-ae00-244319eeabba · outbound

This paper cites Finite-sample analysis of distributed q-learning for multi-agent networks,.

Variance-Reduced Q-Learning over Static and Time-Varying Networks Finite-sample analysis of distributed q-learning for multi-agent networks,

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-01T06:37:06.414302Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T06:37:06.414302Z digest=sha256:de31478a2bad2e29a66e4b84df6d1aa12e3eaf47ed3fb66903251e7a7f35ddbb

Observation eee3fe86-1ea7-4daf-be87-9f0d77183c7b · outbound

This paper cites A finite-time analysis of distributed Q- Learning,.

Variance-Reduced Q-Learning over Static and Time-Varying Networks A finite-time analysis of distributed Q- Learning,

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-01T06:37:06.433768Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T06:37:06.433768Z digest=sha256:89a889da5365e30b0380ea8231dccf52206365a08111e48061a1f76534ff2177

Observation 98745f85-674a-40f2-8387-410c49443fdb · outbound

This paper cites Finite-time convergence rates of decentralized stochastic approximation with applications in multi-agent and multi-task learning,.

Variance-Reduced Q-Learning over Static and Time-Varying Networks Finite-time convergence rates of decentralized stochastic approximation with applications in multi-agent and multi-task learning,

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-01T06:37:06.445062Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T06:37:06.445062Z digest=sha256:55ce88f039012e988a178930d9de371a853a0177aebbc7e42a74c3e553ce9d05

Observation 014820b0-3445-4a60-810d-bbd9173529ec · outbound

This paper cites Federated reinforcement learning: Linear speedup under Markovian sampling,.

Variance-Reduced Q-Learning over Static and Time-Varying Networks Federated reinforcement learning: Linear speedup under Markovian sampling,

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-01T06:37:06.456743Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T06:37:06.456743Z digest=sha256:1c278f3c5a785ef24f37615348da58056ca81d1184499d165ec059e87eec3ff1

Observation 65aec993-d6c6-46b8-8e1e-5299178c5323 · outbound

This paper cites The blessing of heterogeneity in federated Q-learning: Linear speedup and beyond,.

Variance-Reduced Q-Learning over Static and Time-Varying Networks The blessing of heterogeneity in federated Q-learning: Linear speedup and beyond,

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-01T06:37:06.470700Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T06:37:06.470700Z digest=sha256:20c530afaefc437ef5011afd4f5922c5ae748afbf490d9aee0cccb733c97a1ba

Observation 2cae0d64-3db1-4760-bb3d-635b082d4cf8 · outbound

This paper cites Federated Temporal Difference Learning with Linear Function Approximation under Environmental Heterogeneity.

Variance-Reduced Q-Learning over Static and Time-Varying Networks Federated Temporal Difference Learning with Linear Function Approximation under Environmental Heterogeneity

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-01T06:37:06.478589Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T06:37:06.478589Z digest=sha256:018d0aac0c867886e4fdb93aff8c438fb784a729c45b8b05a0bc974b40648104

Observation 10499641-dee2-4774-8b20-cd75bd59f596 · outbound

This paper cites Stochastic approximation with cone-contractive operators: Sharp $\ell_\infty$-bounds for $Q$-learning.

Variance-Reduced Q-Learning over Static and Time-Varying Networks Stochastic approximation with cone-contractive operators: Sharp $\ell_\infty$-bounds for $Q$-learning

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-01T06:37:06.486395Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T06:37:06.486395Z digest=sha256:9868a5cdae9e538d22c86912b550a6e38f77246aca1357b58da4f47ee36be2a4

Observation 161bdcfe-7f7c-41e3-8dcf-95f2dceaa112 · outbound

This paper cites Finite-time analysis of asynchronous stochastic approximation and Q-learning,.

Variance-Reduced Q-Learning over Static and Time-Varying Networks Finite-time analysis of asynchronous stochastic approximation and Q-learning,

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-01T06:37:06.497436Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T06:37:06.497436Z digest=sha256:fff15615d47290a1c3477dbd3a7fdca887cb141346a6966130f70b95fceeae99

Observation 0f76ec8e-0f79-4292-b471-4ab57ac69d48 · outbound

This paper cites Is Q-learning minimax optimal? a tight sample complexity analysis,.

Variance-Reduced Q-Learning over Static and Time-Varying Networks Is Q-learning minimax optimal? a tight sample complexity analysis,

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-01T06:37:06.504912Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T06:37:06.504912Z digest=sha256:a2ec921e891cadd70c7f3e190cf504ab98b6988073d79765524e7cbe81b6d9ae

Observation 1ffea82b-1250-4fda-bca2-2f3ae5fe7b4b · outbound

This paper cites an unresolved cited work.

Variance-Reduced Q-Learning over Static and Time-Varying Networks Unresolved cited work

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-01T06:37:06.510819Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T06:37:06.510819Z digest=sha256:20f946dfdbfbee92e218ff5bf6949afeed4dffd83c23dafac7e2a92d0cd30303

Observation 14d470a4-ad2f-452f-a964-94841bef1a45 · outbound

This paper cites Finite-sample convergence rates for Q- learning and indirect algorithms,.

Variance-Reduced Q-Learning over Static and Time-Varying Networks Finite-sample convergence rates for Q- learning and indirect algorithms,

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-01T06:37:06.528957Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T06:37:06.528957Z digest=sha256:1055d87f2bb4950aa22394a5d95e8681ce4b230fa07bcd69d265a5106c04414c

Observation e0caff71-cb76-4246-8ec9-ad4b06d5c62b · outbound

This paper cites Near-optimal time and sample complexities for solving Markov decision processes with a generative model,.

Variance-Reduced Q-Learning over Static and Time-Varying Networks Near-optimal time and sample complexities for solving Markov decision processes with a generative model,

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-01T06:37:06.535030Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T06:37:06.535030Z digest=sha256:66dc50095530e08820c4f4a816c7230858158c6b34896959b1a80797b319ec17

Observation 7a135b4b-075a-47e6-8a7e-4712901c3751 · outbound

This paper cites Asynchronous stochastic approximation and Q- learning,.

Variance-Reduced Q-Learning over Static and Time-Varying Networks Asynchronous stochastic approximation and Q- learning,

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-01T06:37:06.544387Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T06:37:06.544387Z digest=sha256:ba6d02afc58a9fef8c8953145ac7f7cb24c10a62ad3c9e68cb6a35cd16cca9c2

Observation 28c7f9d6-4a4c-41af-8a6c-4ae7918ea907 · outbound

This paper cites Achieving geometric conver- gence for distributed optimization over time-varying graphs,.

Variance-Reduced Q-Learning over Static and Time-Varying Networks Achieving geometric conver- gence for distributed optimization over time-varying graphs,

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-01T06:37:06.550135Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T06:37:06.550135Z digest=sha256:675d37c9006f56de14d7f75c000ca71d93730edad5e5df8c90c1b233c5637197

Observation 806a3594-2d35-4efe-a117-f7a32443d8f9 · outbound

This paper cites High-Dimensional Statistics.

Variance-Reduced Q-Learning over Static and Time-Varying Networks High-Dimensional Statistics

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-01T06:37:06.555728Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T06:37:06.555728Z digest=sha256:b197dcc0f6106e3c4f5916182c98edaedf105f02c7d5cf435b16ddb978d39c24

Observation 75451653-c788-44ea-a154-8a5240078d81 · outbound

This paper cites Lattimore and C.

Variance-Reduced Q-Learning over Static and Time-Varying Networks Lattimore and C

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-01T06:37:06.562214Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T06:37:06.562214Z digest=sha256:f097d71cf0db18893e692a4de6a758c9d52575b9f9754c62cd17a8a9d35ddc54

Pith citing papers

No inbound Pith citation observations are available.