Pith. sign in

Paper Citation Record · LEDGER

Convergence Analysis of the Last Iterate in Distributed Stochastic Gradient Descent with Momentum

As of 19 August 2026, this Paper Citation Record lists 38 of 38 outbound references and 0 inbound Pith citation observations for arXiv:2505.10889.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.10889 v1

Coverage vector

measured 38 of 38 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-15T21:16:05.092207Z

measured 38 of 38 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-18T06:34:40.430872+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

38 of 38 outbound references displayed

  • verified exact0
  • verified fuzzy30
  • unresolved8
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 1b5f86fc-1484-450f-b64d-78e01a15e9ff · outbound

This paper cites A stochastic approximation method,.

Convergence Analysis of the Last Iterate in Distributed Stochastic Gradient Descent with Momentum A stochastic approximation method,

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-15T21:16:04.870210Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:16:04.870210Z digest=sha256:321870567d90eea49284f595d1c7adccc7c76e3406196c05213e0e22aaadb8d4

Observation 7585abd5-81d3-41f0-8ca3-fcf92f29d43f · outbound

This paper cites Some methods of speeding up the convergence of iteration methods,.

Convergence Analysis of the Last Iterate in Distributed Stochastic Gradient Descent with Momentum Some methods of speeding up the convergence of iteration methods,

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:16:05.816393Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T21:16:04.875811Z digest=sha256:369b8cba6293d7703ae098f8787de0bea96eab2f70bf5399f9f8c933dc8dd334

Observation b807a7cb-5537-4f4d-a557-2a99fad654c7 · outbound

This paper cites Adaptive deep feature learning network with Nesterov momentum and its application to rotating machinery fault diagnosis,.

Convergence Analysis of the Last Iterate in Distributed Stochastic Gradient Descent with Momentum Adaptive deep feature learning network with Nesterov momentum and its application to rotating machinery fault diagnosis,

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:16:05.799392Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T21:16:04.881935Z digest=sha256:a7e3f8c1a899b9b87de84381b9035f2ff1ff40d7bdc5d5dedff09ed633c40a99

Observation 9e5d4f43-4c8a-4c92-a6fc-8b05117e671b · outbound

This paper cites Combining ordered subsets and momentum for accelerated X-ray CT image reconstruction,.

Convergence Analysis of the Last Iterate in Distributed Stochastic Gradient Descent with Momentum Combining ordered subsets and momentum for accelerated X-ray CT image reconstruction,

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:16:05.783110Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T21:16:04.887847Z digest=sha256:4656b2884dbbef907cd64212065b59bb367c13ebcb228bf76b8be0a4423818ec

Observation d25caf16-974f-4e6e-9b89-42683cde325b · outbound

This paper cites Speech recognition with deep recurrent neural networks,.

Convergence Analysis of the Last Iterate in Distributed Stochastic Gradient Descent with Momentum Speech recognition with deep recurrent neural networks,

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:16:05.764791Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T21:16:04.900026Z digest=sha256:8315650b4ff976bf7086726dc64e9300c4000a1140015779ad6154ac4413411a

Observation 2311f3d8-5143-4f18-ac09-c0031f7e3d24 · outbound

This paper cites SGD and Hogwild! convergence without the bounded gradients assumption,.

Convergence Analysis of the Last Iterate in Distributed Stochastic Gradient Descent with Momentum SGD and Hogwild! convergence without the bounded gradients assumption,

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:16:05.745177Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T21:16:04.909718Z digest=sha256:fe931de09487080a6ab7cdd9ef4de71379e3d9fb837e1678b77a595e66213cad

Observation 61dccd6c-5dc3-4985-a3e2-482351608cf1 · outbound

This paper cites Reducing the dimensionality of data with neural networks,.

Convergence Analysis of the Last Iterate in Distributed Stochastic Gradient Descent with Momentum Reducing the dimensionality of data with neural networks,

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-15T21:16:04.917004Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:16:04.917004Z digest=sha256:3287c96529ea7c872c4d077e643a0b6949d9edc591d8b1c4df002fac1655b077

Observation 28293377-7842-44e8-bbee-1a415bf18e64 · outbound

This paper cites Distributed training strategies for the structured perceptron,.

Convergence Analysis of the Last Iterate in Distributed Stochastic Gradient Descent with Momentum Distributed training strategies for the structured perceptron,

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:16:05.712772Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T21:16:04.923055Z digest=sha256:f1e253b5abe635cdc6ba39cb2c740db6a612548aeec2f9a8db3165e2ecaa8d3f

Observation 09b2f917-3779-4625-9eea-956a35f780f4 · outbound

This paper cites Deep learning with elastic averaging sgd,.

Convergence Analysis of the Last Iterate in Distributed Stochastic Gradient Descent with Momentum Deep learning with elastic averaging sgd,

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:16:05.694937Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T21:16:04.928793Z digest=sha256:fc384d0de6ba6883fc911f1d7c726479392ed4fccaadb9ae2dba28b82a846dd2

Observation be118f36-a1fd-4cd2-adfc-7a1d16366ecf · outbound

This paper cites Network topology and communication-computation tradeoffs in decentralized optimization,.

Convergence Analysis of the Last Iterate in Distributed Stochastic Gradient Descent with Momentum Network topology and communication-computation tradeoffs in decentralized optimization,

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:16:05.677857Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T21:16:04.934245Z digest=sha256:3feb2dd22cb646f2ca244e313bd6c8b3cda5208826ebe244c9d524d1b1b9a8e8

Observation 8ba79a71-f9ed-4ec2-b28a-5e96becd7255 · outbound

This paper cites Local SGD Converges Fast and Communicates Little.

Convergence Analysis of the Last Iterate in Distributed Stochastic Gradient Descent with Momentum Local SGD Converges Fast and Communicates Little

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-15T21:16:04.939109Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:16:04.939109Z digest=sha256:d37a35f09770eb95292617f3f21b6a7aeb834ac062034773fef29a4bdfdfdbea

Observation 3c49b6d1-4c9e-4dce-994f-4bbc64568414 · outbound

This paper cites Parallel restarted sgd with faster convergence and less communication: Demystifying why model averaging works for deep learning,.

Convergence Analysis of the Last Iterate in Distributed Stochastic Gradient Descent with Momentum Parallel restarted sgd with faster convergence and less communication: Demystifying why model averaging works for deep learning,

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:16:05.658384Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T21:16:04.945190Z digest=sha256:dc55fe7411b98523e30782d49070d734e469399cc93fdbb3fc36ae62d852a0fa

Observation e04901f1-2fae-44e9-86a5-c19c7cef8610 · outbound

This paper cites A linear speedup analysis of distributed deep learning with sparse and quantized communication,.

Convergence Analysis of the Last Iterate in Distributed Stochastic Gradient Descent with Momentum A linear speedup analysis of distributed deep learning with sparse and quantized communication,

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:16:05.638473Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T21:16:04.950550Z digest=sha256:bb8e25f09e81027f72923ecb7d8ee1a3ef955d7bcbf2fbdc42bfb699d06047f9

Observation d7cc35b9-7ac1-4333-b872-59d8a830cf17 · outbound

This paper cites Can decentralized algorithms outperform centralized algorithms? a case study for decentralized parallel stochastic gradient descent,.

Convergence Analysis of the Last Iterate in Distributed Stochastic Gradient Descent with Momentum Can decentralized algorithms outperform centralized algorithms? a case study for decentralized parallel stochastic gradient descent,

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-15T21:16:04.959122Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:16:04.959122Z digest=sha256:b3beede8eeaa3e694ecc614338c3c34bd2be2e6b6893dab776895186f217a910

Observation bd5fc640-7e5a-4828-8cea-377fc323f43a · outbound

This paper cites Collaborative deep learning in fixed topology networks,.

Convergence Analysis of the Last Iterate in Distributed Stochastic Gradient Descent with Momentum Collaborative deep learning in fixed topology networks,

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:16:05.609126Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T21:16:04.965081Z digest=sha256:6fec248e6fad96c8901d1ef7a341c33122aa695d1facacd2cbcc27136ad39d54

Observation 54d6075a-44b8-4aa7-89e9-dae226ecf6aa · outbound

This paper cites On nonconvex decentralized gradient descent,.

Convergence Analysis of the Last Iterate in Distributed Stochastic Gradient Descent with Momentum On nonconvex decentralized gradient descent,

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-15T21:16:04.970086Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:16:04.970086Z digest=sha256:4537a937029d2c737b8924478db30e52fdc94d20de8bd00081e74bbc04fba06b

Observation 21478628-8e38-4c4f-bc5a-787d95886490 · outbound

This paper cites Cooperative sgd: A unified framework for the design and analysis of local-update sgd algorithms,.

Convergence Analysis of the Last Iterate in Distributed Stochastic Gradient Descent with Momentum Cooperative sgd: A unified framework for the design and analysis of local-update sgd algorithms,

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:16:05.578632Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T21:16:04.974941Z digest=sha256:3e74fae66e06b6a790ac8ab3405b4277e6af81d70d0e39907e01fb84e8617d8e

Observation 5968c1dd-000c-4260-85ef-567626f646cd · outbound

This paper cites Imagenet classification with deep convolutional neural networks,.

Convergence Analysis of the Last Iterate in Distributed Stochastic Gradient Descent with Momentum Imagenet classification with deep convolutional neural networks,

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:16:05.560596Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T21:16:04.981179Z digest=sha256:1ba780c30e20c773b1e13458d33f7eeeee6d1ceca58ed3a892cdd31e96121f82

Observation 063a5a98-a5ae-47e7-bb83-de7c48a3e4b8 · outbound

This paper cites A Unified Analysis of Stochastic Momentum Methods for Deep Learning.

Convergence Analysis of the Last Iterate in Distributed Stochastic Gradient Descent with Momentum A Unified Analysis of Stochastic Momentum Methods for Deep Learning

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-15T21:16:04.987777Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:16:04.987777Z digest=sha256:c1e65ee0d78d54212161d37d551d1306cd9a1890d0a491ea741eaf8ee91cfcd3

Observation 22d8a116-b5a8-472d-827d-3cd5064a03e6 · outbound

This paper cites On the importance of initialization and momentum in deep learning,.

Convergence Analysis of the Last Iterate in Distributed Stochastic Gradient Descent with Momentum On the importance of initialization and momentum in deep learning,

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:16:05.542612Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T21:16:04.995072Z digest=sha256:e9b6c17904968b7f6bdfb01cee777b992a91c9323d3c33309158e4fb39c629b3

Observation d3c4ab34-6398-4381-a1d0-c753a7476d50 · outbound

This paper cites On the linear speedup analysis of communication efficient momentum sgd for distributed non- convex optimization,.

Convergence Analysis of the Last Iterate in Distributed Stochastic Gradient Descent with Momentum On the linear speedup analysis of communication efficient momentum sgd for distributed non- convex optimization,

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:16:05.524880Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T21:16:05.001300Z digest=sha256:205c7faf58117d63281950ad18f2f851272ec1ce436ee8ebeb3299c5f1408e8e

Observation 615cb8d7-1509-4876-838e-b19ce8728ba0 · outbound

This paper cites On consensus- optimality trade-offs in collaborative deep learning,.

Convergence Analysis of the Last Iterate in Distributed Stochastic Gradient Descent with Momentum On consensus- optimality trade-offs in collaborative deep learning,

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:16:05.507346Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T21:16:05.006388Z digest=sha256:d0e06d22f949e63019146ba8c55b35ea4e47a53cbcef16c6c20ffde3d1877681

Observation 3be9fce3-33cf-4350-af62-f99cc5b9cbdc · outbound

This paper cites Deep gradient compression: Reducing the communication bandwidth for distributed training,.

Convergence Analysis of the Last Iterate in Distributed Stochastic Gradient Descent with Momentum Deep gradient compression: Reducing the communication bandwidth for distributed training,

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:16:05.484749Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T21:16:05.011429Z digest=sha256:3ea7cb983e0cbf4a902eef1618db542ce99138a39d0435c0910a982d237c6e5a

Observation 351bc966-cd6c-4988-b96b-29baf7d61ca0 · outbound

This paper cites Can decentralized algorithms outperform centralized algorithms? a case study for decentralized parallel stochastic gradient descent,.

Convergence Analysis of the Last Iterate in Distributed Stochastic Gradient Descent with Momentum Can decentralized algorithms outperform centralized algorithms? a case study for decentralized parallel stochastic gradient descent,

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:16:05.463015Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T21:16:05.016562Z digest=sha256:43fb918115769dda2b57decf4ed4b5b93cca52445847d425798b07e8e6f3cc78

Observation 2af36f5a-24eb-46ea-b0fa-ca8a190c034b · outbound

This paper cites Decentlam: Decentralized momentum sgd for large-batch deep training,.

Convergence Analysis of the Last Iterate in Distributed Stochastic Gradient Descent with Momentum Decentlam: Decentralized momentum sgd for large-batch deep training,

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:16:05.445325Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T21:16:05.022135Z digest=sha256:ab0da856bafb85bbe26ce12fe904e83f701caf879819a0dfe518b9a843277e14

Observation 800f32fa-3c7f-41c4-a694-f859fd295700 · outbound

This paper cites 2020SQuARM: Communication-efficient momentum SGD for decentralized optimization,.

Convergence Analysis of the Last Iterate in Distributed Stochastic Gradient Descent with Momentum 2020SQuARM: Communication-efficient momentum SGD for decentralized optimization,

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:16:05.422343Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T21:16:05.026972Z digest=sha256:54433581a7d7a83046703bfdc5291f66e7e81ca702476938c25286ce6401c604

Observation 03643d56-567b-42d2-bde3-7f2e2674825f · outbound

This paper cites Periodic Stochastic Gradient Descent with Momentum for Decentralized Training.

Convergence Analysis of the Last Iterate in Distributed Stochastic Gradient Descent with Momentum Periodic Stochastic Gradient Descent with Momentum for Decentralized Training

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-15T21:16:05.031704Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:16:05.031704Z digest=sha256:4a457724b5e93c33b2348b5e2a54c01a5425375578c220d7f3065722bd577c89

Observation e6b9600d-3953-418e-a5ae-5b625c772bdc · outbound

This paper cites Decentralized deep learning using momentum-accelerated consensus,.

Convergence Analysis of the Last Iterate in Distributed Stochastic Gradient Descent with Momentum Decentralized deep learning using momentum-accelerated consensus,

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:16:05.403893Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T21:16:05.037280Z digest=sha256:81a24312b258736061c8f6ba79f76ab321ba4703ba33c24592c47c424bcf3982

Observation 2039e531-6c49-4595-b67b-0fedeebd52be · outbound

This paper cites Parallel restarted sgd with faster convergence and less communication: Demystifying why model averaging works for deep learning,.

Convergence Analysis of the Last Iterate in Distributed Stochastic Gradient Descent with Momentum Parallel restarted sgd with faster convergence and less communication: Demystifying why model averaging works for deep learning,

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:16:05.384740Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T21:16:05.042769Z digest=sha256:53e7fb7355f6e51b55fd34e377f578cb3294ce556b706f24634e9f00dcbc786a

Observation 58a2e080-9d1c-455e-84a1-8ad434aad158 · outbound

This paper cites On the convergence of mSGD and AdaGrad for stochastic optimization,.

Convergence Analysis of the Last Iterate in Distributed Stochastic Gradient Descent with Momentum On the convergence of mSGD and AdaGrad for stochastic optimization,

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:16:05.365253Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T21:16:05.047649Z digest=sha256:9cec300454a6c457b33ed666cb76fe69a2058a58ec745db3b81a19bc7af21a9e

Observation 4cbe6bf7-67b3-45b1-bdbf-c4747bbdeefc · outbound

This paper cites Don't Decay the Learning Rate, Increase the Batch Size.

Convergence Analysis of the Last Iterate in Distributed Stochastic Gradient Descent with Momentum Don't Decay the Learning Rate, Increase the Batch Size

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-15T21:16:05.053234Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:16:05.053234Z digest=sha256:7f46ff827cc2e9ec6af34a82ee3f2dd246f344c922b8512672371c497f03cf1e

Observation 696e5d03-2482-4598-9f3f-27ea1dba391d · outbound

This paper cites Bayesian learning via stochastic gradient langevin dynamics,.

Convergence Analysis of the Last Iterate in Distributed Stochastic Gradient Descent with Momentum Bayesian learning via stochastic gradient langevin dynamics,

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:16:05.346931Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T21:16:05.059295Z digest=sha256:3e75121a4413e2e83f2769595e3a0da6b3ade9503a7ca8bba6a3fe992ac40795

Observation cfc784d5-d69d-4e7c-9cc0-f008fbdc5273 · outbound

This paper cites Convergence of proximal-gradient stochastic variational inference under non-decreasing step-size sequence,.

Convergence Analysis of the Last Iterate in Distributed Stochastic Gradient Descent with Momentum Convergence of proximal-gradient stochastic variational inference under non-decreasing step-size sequence,

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:16:05.328386Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T21:16:05.065059Z digest=sha256:7c4f16c8295fc2bd9b251ed39495e216e3642aed9c468ad68d7df82495005c5e

Observation d656d89c-cd16-4138-b152-92cb33613746 · outbound

This paper cites Understanding the role of momentum in stochastic gradient methods,.

Convergence Analysis of the Last Iterate in Distributed Stochastic Gradient Descent with Momentum Understanding the role of momentum in stochastic gradient methods,

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:16:05.308628Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T21:16:05.070661Z digest=sha256:d6c93283a253732dc8411ce12c7cdeff4920db9cec5931a988d4e793033f3da2

Observation a5415d8d-8810-476f-b99f-1868005ba493 · outbound

This paper cites Deep residual learning for image recognition,.

Convergence Analysis of the Last Iterate in Distributed Stochastic Gradient Descent with Momentum Deep residual learning for image recognition,

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:16:05.288603Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T21:16:05.075459Z digest=sha256:d35bc7798f3cbdcf412550fe6a8a8914c6a493a46d6215a992bfeee282fa9709

Observation 0214823a-791f-4ca3-93aa-1578f2ad0091 · outbound

This paper cites Nesterov,Introductory Lectures on Convex Optimization: A Basic Course.

Convergence Analysis of the Last Iterate in Distributed Stochastic Gradient Descent with Momentum Nesterov,Introductory Lectures on Convex Optimization: A Basic Course

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:16:05.266884Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T21:16:05.081368Z digest=sha256:a3df0358075993a7b45b2170f1f37effd129e43af380a1b05e274b8b01027161

Observation e8d7a3ba-bd4b-439c-8769-b5e6561f3fe8 · outbound

This paper cites Revisit last- iterate convergence of msgd under milder requirement on step size.

Convergence Analysis of the Last Iterate in Distributed Stochastic Gradient Descent with Momentum Revisit last- iterate convergence of msgd under milder requirement on step size

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:16:05.245733Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T21:16:05.086621Z digest=sha256:44fc26be955627c312f0b0d6a57b4c8aeeab979e2040d695d406c61030eea2f8

Observation 4d126e17-d5ef-42e8-9b5e-e2e73f0ad451 · outbound

This paper cites On almost sure convergence for sums of stochastic sequence,.

Convergence Analysis of the Last Iterate in Distributed Stochastic Gradient Descent with Momentum On almost sure convergence for sums of stochastic sequence,

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:16:05.225176Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T21:16:05.092207Z digest=sha256:440a88cb38cceb3819a9aab47e1dc73a891cbfc1be45cd002566b8edabf60ec8

Pith citing papers

No inbound Pith citation observations are available.