Pith. sign in

Paper Citation Record · LEDGER

Sharp higher order convergence rates for the Adam optimizer

As of 17 August 2026, this Paper Citation Record lists 41 of 41 outbound references and 4 inbound Pith citation observations for arXiv:2504.19426.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2504.19426 v1

Coverage vector

measured 41 of 41 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-16T06:04:09.790187Z

measured 45 of 45 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-17T06:30:58.91139+00:00

measured 4 of 4 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T10:42:54.670505Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-21T10:30:00.239783Z

Reference resolution

41 of 41 outbound references displayed

  • verified exact1
  • verified fuzzy19
  • unresolved21
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 1e545ef3-3bc4-48e7-8ca9-5d27c75a4428 · outbound

This paper cites Learning Theory from First Principles , first edition ed.

Sharp higher order convergence rates for the Adam optimizer Learning Theory from First Principles , first edition ed

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T06:04:10.381372Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-16T06:04:09.610161Z digest=sha256:da97cf42f8dce04dd4b2aa9fc9e1b9925997aa0bd52c69045fa720fda11b54f6

Observation 95094f53-0767-4ba4-9e0b-96382f5163ea · outbound

This paper cites Convergence Analysis of a Momentum Algorithm with Adaptive Step Size for Non Convex Optimization.

Sharp higher order convergence rates for the Adam optimizer Convergence Analysis of a Momentum Algorithm with Adaptive Step Size for Non Convex Optimization

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-16T06:04:09.614937Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T06:04:09.614937Z digest=sha256:48abf96422933a93b1201c6da835b93cf6358f9a558f5a39327d6c815a6281cb

Observation d3473149-1eb9-46ed-8deb-649ed03aefc0 · outbound

This paper cites Convergence and dynamical behavior of the Adam algorithm for nonconvex stochastic optimization.

Sharp higher order convergence rates for the Adam optimizer Convergence and dynamical behavior of the Adam algorithm for nonconvex stochastic optimization

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T06:04:10.367354Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-16T06:04:09.620337Z digest=sha256:b6960f30c546a57719c5b48bded3dc8b87a8ce9884981e3f8c6cefc2e8c77dba

Observation 4977aa62-6d09-4f11-98c2-6c0ce428ce48 · outbound

This paper cites Stochastic optimization with momentum: convergence, fluctuations, an d traps avoidance.

Sharp higher order convergence rates for the Adam optimizer Stochastic optimization with momentum: convergence, fluctuations, an d traps avoidance

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T06:04:10.353373Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-16T06:04:09.624754Z digest=sha256:7abb51a6e85bfa761d19de4351f1afbc5844484966138d49d5a061390d5e8636

Observation 0510455f-8c72-441d-b427-9b81a1f9c74a · outbound

This paper cites S., and von Wurstemberge r, P.

Sharp higher order convergence rates for the Adam optimizer S., and von Wurstemberge r, P

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T06:04:10.339483Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-16T06:04:09.629160Z digest=sha256:17497360f7cc1d7b6f69de46e01fef652804912673aedef1e7b1ac363b2146bd

Observation c226eefd-014e-434d-a565-0bad11a6b358 · outbound

This paper cites A Proof of Local Convergence for the Adam Optimizer.

Sharp higher order convergence rates for the Adam optimizer A Proof of Local Convergence for the Adam Optimizer

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T06:04:10.325588Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-16T06:04:09.633633Z digest=sha256:01ed3f896ad691a7d53004a7b558e989327eb5aec6d1c442f17296d081783a2c

Observation f410d068-3ece-4c72-98bb-22e205b5344f · outbound

This paper cites Méthode générale pour la résolution des systèmes d’équatio ns si- multanées.

Sharp higher order convergence rates for the Adam optimizer Méthode générale pour la résolution des systèmes d’équatio ns si- multanées

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T06:04:10.311695Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-16T06:04:09.638642Z digest=sha256:acb30b5668c844d5351a93af89a8d05bef60578dd13eaa78882c61e9491f6083

Observation d04069a4-3cbd-4265-bc64-8e5f25673c14 · outbound

This paper cites On the Convergence of A Class of Adam-Type Algorithms for Non-Convex Optimization.

Sharp higher order convergence rates for the Adam optimizer On the Convergence of A Class of Adam-Type Algorithms for Non-Convex Optimization

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-16T06:04:09.643067Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T06:04:09.643067Z digest=sha256:07b8ee39021ada4baac87f03c740ea1776fa175b282e39c72caf68b9613a7651

Observation 5c769ceb-4226-4f82-980b-d63e70b26a80 · outbound

This paper cites Non-convergence of stochas- tic gradient descent in the training of deep neural networks.

Sharp higher order convergence rates for the Adam optimizer Non-convergence of stochas- tic gradient descent in the training of deep neural networks

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T06:04:10.298432Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-16T06:04:09.647659Z digest=sha256:5953955096eeefc4970be4f2fbaee62d98168c88739dc33625305c699bc93310

Observation 78f34057-0ea6-4265-b2e4-ee8d9fc8e44f · outbound

This paper cites A Simple Convergence Proof of Adam and Adagrad.

Sharp higher order convergence rates for the Adam optimizer A Simple Convergence Proof of Adam and Adagrad

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T06:04:10.285128Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-16T06:04:09.652016Z digest=sha256:d55890e0cd1fcf80f266fc627e2ec98f69210d71db62d021383f1eb80e36c807

Observation 92dd69be-8d6a-467e-a83f-5d3eb6327a2a · outbound

This paper cites Non-convergence of Adam and other adaptive stochastic gradient descent optimization methods for non-vanishing learning rates.

Sharp higher order convergence rates for the Adam optimizer Non-convergence of Adam and other adaptive stochastic gradient descent optimization methods for non-vanishing learning rates

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-16T06:04:09.656396Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T06:04:09.656396Z digest=sha256:7f7afb57599ec4abfba482db0dc4073ac8d4a8b7e077a308e0e68fbb786b3ffc

Observation 58005b27-ec72-479c-9bf0-f06098d20999 · outbound

This paper cites Convergence rates for the Adam optimizer.

Sharp higher order convergence rates for the Adam optimizer Convergence rates for the Adam optimizer

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-16T06:04:09.661231Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T06:04:09.661231Z digest=sha256:7c90754850ac80e6182939def2a487a7177c4ca061d0e9e6a268d7305b195b7b

Observation 30dfd595-5127-4ed5-8461-c3e428b208bc · outbound

This paper cites Learning rate adaptive stochastic gradient descent optimization methods: numerical simulations for deep learning methods for partial differential equations and convergence analyses.

Sharp higher order convergence rates for the Adam optimizer Learning rate adaptive stochastic gradient descent optimization methods: numerical simulations for deep learning methods for partial differential equations and convergence analyses

Reference 13

Resolution
verified exact
local_arxiv, observed 2026-08-16T06:04:10.062227Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-16T06:04:09.665987Z digest=sha256:91cc598802e760c1a5a6baba16e07eddaa8d89bfe8f3d2b7f0e25d68545a7a63

Observation fc0e604e-8ec4-4439-9be1-c37486970d8c · outbound

This paper cites Convergence of stochastic gradient descent schemes for Lojasiewicz-landscapes.

Sharp higher order convergence rates for the Adam optimizer Convergence of stochastic gradient descent schemes for Lojasiewicz-landscapes

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-16T06:04:09.670445Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T06:04:09.670445Z digest=sha256:24b88f88aa5d680c067de0f7db663f2aa2701c0ccd83deb07d9e3300e612a90d

Observation 4838e261-417c-4152-be2f-a0ab65616662 · outbound

This paper cites Non-convergence to global minimizers in data driven supervised deep learning: Adam and stochastic gradient descent optimization provably fail to converge to global minimizers in the training of deep neural networks with ReLU activation.

Sharp higher order convergence rates for the Adam optimizer Non-convergence to global minimizers in data driven supervised deep learning: Adam and stochastic gradient descent optimization provably fail to converge to global minimizers in the training of deep neural networks with ReLU activation

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-16T06:04:09.675097Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T06:04:09.675097Z digest=sha256:b642fd90eb03c537282b8ef3ca875ae21d1d9f3d1c6a05dd128c681a1da5f18f

Observation 91ca25ad-86d3-4c1e-bf20-8e18f616897e · outbound

This paper cites On the oracle complexity of smooth strongly convex minimization.

Sharp higher order convergence rates for the Adam optimizer On the oracle complexity of smooth strongly convex minimization

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T06:04:10.272092Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-16T06:04:09.679456Z digest=sha256:c43dea935fa34248a9625bde4628cf539ae9f6f285a5d5f032823b2d8025e190

Observation 6a5ec2e9-e7cd-4cb5-87a2-067fcb5b66fd · outbound

This paper cites Adaptive subgradient methods for online learning and stochastic optimization.

Sharp higher order convergence rates for the Adam optimizer Adaptive subgradient methods for online learning and stochastic optimization

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T06:04:10.258671Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-16T06:04:09.683879Z digest=sha256:d62a0f6cd361814536c8a154452726e72a233477c7ce5173cda8f757e4f5a365

Observation b29572d0-e2c8-499e-a84e-d5b69df684f6 · outbound

This paper cites Towards a Mathematical Understanding of Neural Network-Based Machine Learning: what we know and what we don't.

Sharp higher order convergence rates for the Adam optimizer Towards a Mathematical Understanding of Neural Network-Based Machine Learning: what we know and what we don't

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-16T06:04:09.688177Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T06:04:09.688177Z digest=sha256:fd2c1cfb33dbf9d23eaf68f5a35e3cb5f49c7b79a83f0d2824e199b44b858e73

Observation 16cf2906-5a97-40e5-b95f-01a96f2bc6ac · outbound

This paper cites Blow up phenomena for gradient descent optimization methods in the training of artificial neural networks.

Sharp higher order convergence rates for the Adam optimizer Blow up phenomena for gradient descent optimization methods in the training of artificial neural networks

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-16T06:04:09.692953Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T06:04:09.692953Z digest=sha256:649b79fabdd537037224d87f1fe3e99b4659d6e37ecfb971bd464207698274c0

Observation cb1307a7-986c-4c29-a4be-029e263cc436 · outbound

This paper cites Handbook of Convergence Theorems for (Stochastic) Gradient Methods.

Sharp higher order convergence rates for the Adam optimizer Handbook of Convergence Theorems for (Stochastic) Gradient Methods

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-16T06:04:09.697637Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T06:04:09.697637Z digest=sha256:38ff0577f8fd06eb9c1e97c47d1b377d50252a82f8e560dc60598766d3e8a7fc

Observation e9591481-79ff-4299-888f-681aa1e0df85 · outbound

This paper cites Non asymptotic analysis of Adaptive stochastic gradient algorithms and applications.

Sharp higher order convergence rates for the Adam optimizer Non asymptotic analysis of Adaptive stochastic gradient algorithms and applications

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-16T06:04:09.702282Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T06:04:09.702282Z digest=sha256:16e7a2a47c1e85c12d3002a86fad7cb6605ba85998fc7cb634fd71b9d1e4f13f

Observation 449966e4-f16d-4285-a2b5-8a5db4a37bcd · outbound

This paper cites Convergence of Adam for Non-convex Objectives: Relaxed Hyperparameters and Non-ergodic Case.

Sharp higher order convergence rates for the Adam optimizer Convergence of Adam for Non-convex Objectives: Relaxed Hyperparameters and Non-ergodic Case

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-16T06:04:09.706749Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T06:04:09.706749Z digest=sha256:f3c7776914bed1f066d75af8c10e6a184c53d13ac68b360fbfc37954d2dce809

Observation 9d8fc442-befe-4e62-be43-57a196566a6a · outbound

This paper cites Lecture 6e: Rm- sprop: Divide the gradient by a running average of its recent magnitude.

Sharp higher order convergence rates for the Adam optimizer Lecture 6e: Rm- sprop: Divide the gradient by a running average of its recent magnitude

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T06:04:10.245313Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-16T06:04:09.711382Z digest=sha256:b62c42edd3cf98673c355ef8eb9772be52f1c6d8b9d9864584672f47f2af507c

Observation 51d98009-7324-40a1-96fa-0902e09121a4 · outbound

This paper cites Revisiting Convergence of AdaGrad with Relaxed Assumptions.

Sharp higher order convergence rates for the Adam optimizer Revisiting Convergence of AdaGrad with Relaxed Assumptions

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-16T06:04:09.715653Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T06:04:09.715653Z digest=sha256:029e5234f94b36f1ae0801c9fa1261f132338d145e346b13c925a8330aaeff4d

Observation 6dfc630a-a26d-4f45-afa4-ff776ddcc1ac · outbound

This paper cites A., and Johnson, C.

Sharp higher order convergence rates for the Adam optimizer A., and Johnson, C

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T06:04:10.231463Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-16T06:04:09.720111Z digest=sha256:378ba6a2d947629f15ac2ac54c2d955b1c9cab6cf791a7304fa9ba3c6db49e1f

Observation f455d188-0797-4df8-84c2-48014e95f2c1 · outbound

This paper cites Strong error analysis for stochastic gradient descent opti mization algorithms.

Sharp higher order convergence rates for the Adam optimizer Strong error analysis for stochastic gradient descent opti mization algorithms

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T06:04:10.217787Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-16T06:04:09.724439Z digest=sha256:40724e301763181f2ea27b65a2cd64a023945eab71200fece079e67826836997

Observation c281985c-4ec6-4414-b305-3aca83575d19 · outbound

This paper cites Mathematical Introduction to Deep Learning: Methods, Implementations, and Theory.

Sharp higher order convergence rates for the Adam optimizer Mathematical Introduction to Deep Learning: Methods, Implementations, and Theory

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-16T06:04:09.729045Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T06:04:09.729045Z digest=sha256:6ba752bdfb377093a423a4a300b99c61159d32799e57d2de01ecb871497f4851

Observation 19228ced-58a2-4c41-8298-742441154085 · outbound

This paper cites Non-convergence to global minimizers for Adam and stochastic gradient descent optimization and constructions of local minimizers in the training of artificial neural networks.

Sharp higher order convergence rates for the Adam optimizer Non-convergence to global minimizers for Adam and stochastic gradient descent optimization and constructions of local minimizers in the training of artificial neural networks

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-16T06:04:09.733450Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T06:04:09.733450Z digest=sha256:c3d2ef8ad87be43c8767cd5fb4f5f90babfeccd0e8a689e24f4a73e6c83c8a95

Observation a29c38ad-c599-4374-a0af-e36ff323e3c6 · outbound

This paper cites Lower error bounds for the stochas- tic gradient descent optimization algorithm: sharp conver gence rates for slowly and fast decaying learning rates.

Sharp higher order convergence rates for the Adam optimizer Lower error bounds for the stochas- tic gradient descent optimization algorithm: sharp conver gence rates for slowly and fast decaying learning rates

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T06:04:10.202926Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-16T06:04:09.737953Z digest=sha256:15438d9e11a89ea39b9edee62d39e0b0763137ac8ca328ab84962d62e93e4d5e

Observation dae0a935-b83d-4d20-8f31-a0f009c62063 · outbound

This paper cites Adam: A Method for Stochastic Optimization.

Sharp higher order convergence rates for the Adam optimizer Adam: A Method for Stochastic Optimization

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-16T06:04:09.742264Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T06:04:09.742264Z digest=sha256:f5b158da56ed7d7b7575229c027fd88d984698f12eeb89be20e5a098b42e4116

Observation 95ccecfa-b64f-4906-8d25-b0e53182d437 · outbound

This paper cites Convergence of Adam Under Relaxed Assumptions.

Sharp higher order convergence rates for the Adam optimizer Convergence of Adam Under Relaxed Assumptions

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-16T06:04:09.746762Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T06:04:09.746762Z digest=sha256:58864ead7330102d7db9e3fdf6f1c0316056f6268fa9f2485b2a25f38df4e2c7

Observation 86f650ee-260b-426e-8d29-68d3aabf7352 · outbound

This paper cites Dying ReLU and Initialization: Theory and Numerical Exampl es.

Sharp higher order convergence rates for the Adam optimizer Dying ReLU and Initialization: Theory and Numerical Exampl es

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T06:04:10.187787Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-16T06:04:09.751091Z digest=sha256:e0286f811cffb6331f39cf37b5f3901963f702c42078b7507e88b107df28a6f9

Observation 0424415e-49a1-4f4b-8976-8100d8a41505 · outbound

This paper cites A method of solving a convex programming problem with conver - gence rate o(1/k2).

Sharp higher order convergence rates for the Adam optimizer A method of solving a convex programming problem with conver - gence rate o(1/k2)

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T06:04:10.174011Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-16T06:04:09.755626Z digest=sha256:5b7082251c9e3265cdfe537ce34d0269b08a8f752da60e822e1e7ed2f91780aa

Observation fe8a8080-0bbb-41cf-a270-d6e405d79715 · outbound

This paper cites Introductory lectures on convex optimization , vol.

Sharp higher order convergence rates for the Adam optimizer Introductory lectures on convex optimization , vol

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T06:04:10.159993Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-16T06:04:09.759658Z digest=sha256:0f12962e8937128c5ad0a8ac96ddc55867cee417b58737c2677192fd09f3f60e

Observation e24dc468-266b-4b47-b6f0-33c37b1e2b02 · outbound

This paper cites Gradient methods for the minimisation of functionals.

Sharp higher order convergence rates for the Adam optimizer Gradient methods for the minimisation of functionals

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T06:04:10.146073Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-16T06:04:09.764208Z digest=sha256:eace8aaf80dbd41bfc82f591c1b395f76a6f149173113e2deee69caecfbcbe50

Observation 996c2fd0-bf19-4c6e-baca-e2482d5e9ef1 · outbound

This paper cites Some methods of speeding up the convergence of iteration met hods.

Sharp higher order convergence rates for the Adam optimizer Some methods of speeding up the convergence of iteration met hods

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T06:04:10.131640Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-16T06:04:09.768650Z digest=sha256:c28f94588f573bf25cc3d79a0e46dcd5aea16fa5e89ad0ea659fc00c67e2496b

Observation 340a60d1-29f5-4cd8-bf5b-f7156d05fe87 · outbound

This paper cites On the Convergence of Adam and Beyond.

Sharp higher order convergence rates for the Adam optimizer On the Convergence of Adam and Beyond

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-16T06:04:09.772860Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T06:04:09.772860Z digest=sha256:bd10ee8b071ef709a2d9132941ef7d238bda55b8779d877e92bd65656fa973be

Observation 3bf2af90-2139-4dc3-8400-1cf1c3df1241 · outbound

This paper cites An overview of gradient descent optimization algorithms.

Sharp higher order convergence rates for the Adam optimizer An overview of gradient descent optimization algorithms

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-16T06:04:09.777150Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T06:04:09.777150Z digest=sha256:daaa9f6bf2fc9f7a569b125769b0560bf9b9aee637d6e403a7ae37356c7ab269

Observation ade76d91-fa5e-44ae-b3a7-7d77354f4c5e · outbound

This paper cites Optimization for deep learning: theory and algorithms.

Sharp higher order convergence rates for the Adam optimizer Optimization for deep learning: theory and algorithms

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-16T06:04:09.781626Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T06:04:09.781626Z digest=sha256:0c36f6f21a61fce3ba34a2e2977c23d8a0f2e91b33d0b1e975263c486da508e8

Observation 55a86273-5bdb-45fa-b649-7a7b9dfb5551 · outbound

This paper cites Adam Can Converge Without Any Modification On Update Rules.

Sharp higher order convergence rates for the Adam optimizer Adam Can Converge Without Any Modification On Update Rules

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-16T06:04:09.786051Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T06:04:09.786051Z digest=sha256:df4a77b58bf9e7427b35f941275d2b4a2b976872c869d15d69ee66b9df5b865e

Observation 1847c1e1-8652-4047-9445-2fb4f0c4022f · outbound

This paper cites A Sufficient Condition for Convergences of Adam and RMSProp.

Sharp higher order convergence rates for the Adam optimizer A Sufficient Condition for Convergences of Adam and RMSProp

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-16T06:04:09.790187Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T06:04:09.790187Z digest=sha256:287f5e9a4b6f7309bbe27028d300d3024a157427c6a6f0c5206d78f49ec47024

Pith citing papers

Observation 4d82e7a9-7249-454b-a4c3-3b8045202f8a · inbound

Adaptive Preconditioners Trigger Loss Spikes in Adam cites this paper.

Adaptive Preconditioners Trigger Loss Spikes in Adam Sharp higher order convergence rates for the Adam optimizer

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-07T10:42:54.670505Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:42:54.670505Z digest=sha256:2126134326bdafe76607c1fd2be124014338b5ffc8c3cf11537e6282234c9435

Observation 914b32af-e72b-4420-9a8d-1fbd7a790794 · inbound

Global Stability and Step Size Robustness of RMSProp cites this paper.

Global Stability and Step Size Robustness of RMSProp Sharp higher order convergence rates for the Adam optimizer

Reference 11

Resolution
verified exact
arxiv_id, observed 2026-05-21T10:30:00.242328Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-05-21T10:27:38.221096Z digest=sha256:6064f91ae82534d60b146f576ab27aade96c09b5a74649e9283a5e808e177ab2

Observation ce90e50e-8aaa-4330-871e-33521e027eb9 · inbound

Adam-SHANG: A Convergent Adam-Type Method for Stochastic Smooth Convex Optimization cites this paper.

Adam-SHANG: A Convergent Adam-Type Method for Stochastic Smooth Convex Optimization Sharp higher order convergence rates for the Adam optimizer

Reference 13

Resolution
verified exact
arxiv_id, observed 2026-05-14T18:47:36.587701Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-05-14T18:44:25.714500Z digest=sha256:1bb2c4a72cea3dcf6e0bec3bfb703f4404641ff83607e9f812ebaffdf1368040

Observation b5b8a45f-5e63-49d3-9f03-ebbb75c42c96 · inbound

Unified convergence analysis for gradient descent optimization methods in the training of deep neural networks cites this paper.

Unified convergence analysis for gradient descent optimization methods in the training of deep neural networks Sharp higher order convergence rates for the Adam optimizer

Reference 21

Resolution
unresolved
no resolver link, observed 2026-07-11T20:46:05.467029Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-11T20:46:05.467029Z digest=sha256:edcd7421d8fbe63e9faf11f28e8dd43440b293830bc34ebc51084e41c8a268ae