Pith. sign in

Paper Citation Record · LEDGER

PADAM: Parallel averaged Adam reduces the error for stochastic optimization in scientific machine learning

As of 8 August 2026, this Paper Citation Record lists 55 of 55 outbound references and 2 inbound Pith citation observations for arXiv:2505.22085.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.22085 v1

Coverage vector

measured 55 of 55 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T13:20:49.832951Z

measured 57 of 57 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 2 of 2 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-03T10:09:36.912553Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-04T07:29:38.373231Z

Reference resolution

55 of 55 outbound references displayed

  • verified exact4
  • verified fuzzy25
  • unresolved26
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 0ede2bde-50d0-4e6b-b914-b26768d3a5f4 · outbound

This paper cites Adam with model exponential moving average is effective for nonconvex optimization.

PADAM: Parallel averaged Adam reduces the error for stochastic optimization in scientific machine learning Adam with model exponential moving average is effective for nonconvex optimization

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-07T13:20:44.027164Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:20:44.027164Z digest=sha256:f702a7c5e9e17a64717e77fcc155bcd2caecbccd8add85032319541d239ea026

Observation ce1f34b2-c3bc-4452-a082-0e5e2cfea8f5 · outbound

This paper cites General framework for online-to-nonconvex conversion: Schedule-free SGD is also effective for nonconvex optimization.

PADAM: Parallel averaged Adam reduces the error for stochastic optimization in scientific machine learning General framework for online-to-nonconvex conversion: Schedule-free SGD is also effective for nonconvex optimization

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-07T13:20:44.103140Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:20:44.103140Z digest=sha256:711cc1cffa5fd533ab44c7fc8f6b0ad9843532fab62b7a8167d3bbc3a2954e9c

Observation dbc51ba4-6467-4b70-b652-51e18dec4760 · outbound

This paper cites an unresolved cited work.

PADAM: Parallel averaged Adam reduces the error for stochastic optimization in scientific machine learning Unresolved cited work

Reference 3

Resolution
unresolved
raw_fallback, observed 2026-08-07T13:20:59.424343Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T13:20:44.251006Z digest=sha256:d1c7c208f8714f2fbb42ac657c79f164de300e88ec17bf0b3475857380c44eb8

Observation f655e2b0-de84-4ab0-aca7-6b406d49b244 · outbound

This paper cites Learning Theory from First Principles.

PADAM: Parallel averaged Adam reduces the error for stochastic optimization in scientific machine learning Learning Theory from First Principles

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:20:59.101490Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T13:20:44.335953Z digest=sha256:e7f6c776b19c524ea8599d01edace8eff49fc73df71ec9d991c772ad9b50b91d

Observation 003918ea-9e35-4d1e-ac82-3ef4f2fff3e5 · outbound

This paper cites Convergence and dynamical behavior of the Adam algorithm for nonconvex stochastic optimization.

PADAM: Parallel averaged Adam reduces the error for stochastic optimization in scientific machine learning Convergence and dynamical behavior of the Adam algorithm for nonconvex stochastic optimization

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:20:58.683762Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T13:20:44.420377Z digest=sha256:3505f1fd26721c4edb8b4a3e5aa4fd8c1f53aad345df8a59a7e7de8be061f92d

Observation 66f837b2-360b-4c7a-9b98-1a88ab5c6eb6 · outbound

This paper cites Solving the Kolmogorov PDE by means of deep learning.

PADAM: Parallel averaged Adam reduces the error for stochastic optimization in scientific machine learning Solving the Kolmogorov PDE by means of deep learning

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:20:58.425929Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T13:20:44.520870Z digest=sha256:9cecda6164a286b89ce74aaf7729b0525124cb6baf5647b71e5edfe460c2f91e

Observation 603fecab-d84e-4768-8b14-0f59c951ab3d · outbound

This paper cites An overview on deep learning-based approximation methods for partial differential equations.

PADAM: Parallel averaged Adam reduces the error for stochastic optimization in scientific machine learning An overview on deep learning-based approximation methods for partial differential equations

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:20:58.111151Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T13:20:44.587912Z digest=sha256:dbb3ab4132a640f64983d94c06f53f1e3d8521a3fbbb3af3c209e7abfc5c1213

Observation a9c69ac5-8cb3-460f-9215-91f249d6b530 · outbound

This paper cites Solving high-dimensional optimal stopping problems using deep learning.

PADAM: Parallel averaged Adam reduces the error for stochastic optimization in scientific machine learning Solving high-dimensional optimal stopping problems using deep learning

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:20:57.837111Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T13:20:44.675429Z digest=sha256:a5b856631fe6e0dd9f0fee25f58d8a4f4fe70968df8ff0fb79600d5905a76742

Observation fb6639ea-405b-44e7-a370-97f4dd87b3ba · outbound

This paper cites an unresolved cited work.

PADAM: Parallel averaged Adam reduces the error for stochastic optimization in scientific machine learning Unresolved cited work

Reference 9

Resolution
unresolved
raw_fallback, observed 2026-08-07T13:20:57.565534Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T13:20:44.774332Z digest=sha256:2c87a65effb82b8a924e94df5a092516240cad04136567b347892e550f9f73ac

Observation 10695ec8-d505-4f1c-ab3c-475bc9ce6d2b · outbound

This paper cites G., Suau Cuadros, X., and Webb, R.

PADAM: Parallel averaged Adam reduces the error for stochastic optimization in scientific machine learning G., Suau Cuadros, X., and Webb, R

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:20:57.266491Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T13:20:44.911287Z digest=sha256:eb366991470750b312ad3fe3148b1e30b1a9cb5a6d4166d03705ebf62a5c89d6

Observation 0cf9b7d8-f3cb-4c5b-b55e-232ad0c71007 · outbound

This paper cites Scientific machine learning through physics-informed neural networks: where we are and what’s next.

PADAM: Parallel averaged Adam reduces the error for stochastic optimization in scientific machine learning Scientific machine learning through physics-informed neural networks: where we are and what’s next

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:20:56.991187Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T13:20:45.018646Z digest=sha256:35a4f0c555fb09e03d016b531b2fc638e0a3a418aa1098793c6e96562f09c51e

Observation a8dddf37-23af-43f4-b8ac-005545f927e8 · outbound

This paper cites The Road Less Scheduled.

PADAM: Parallel averaged Adam reduces the error for stochastic optimization in scientific machine learning The Road Less Scheduled

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-07T13:20:45.123106Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:20:45.123106Z digest=sha256:0d36a2dd80814e0cb294b2cf410ac34eae892d1defbb7e1e02f9dcdcc2df0c73

Observation 51a48e79-e6d9-435a-9995-498bf0e27a9b · outbound

This paper cites A Simple Convergence Proof of Adam and Adagrad.

PADAM: Parallel averaged Adam reduces the error for stochastic optimization in scientific machine learning A Simple Convergence Proof of Adam and Adagrad

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:20:56.670464Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T13:20:45.226863Z digest=sha256:4ed1e6f00328a0e8d024438476809abeb467266f11881ade1842203ca0a8f32f

Observation c3cb9ab0-2fc9-45ed-ad59-5f53b7904919 · outbound

This paper cites General multilevel adaptations for stochastic approximation algorithms II: CLTs.

PADAM: Parallel averaged Adam reduces the error for stochastic optimization in scientific machine learning General multilevel adaptations for stochastic approximation algorithms II: CLTs

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:20:56.382981Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T13:20:45.364610Z digest=sha256:e0e2186288d3788cf1c54e6d17f474c92eca33822aeb9305010d005ccb934f57

Observation 7561269f-221b-42ab-b3e7-386f24565866 · outbound

This paper cites Convergence rates for the Adam optimizer.

PADAM: Parallel averaged Adam reduces the error for stochastic optimization in scientific machine learning Convergence rates for the Adam optimizer

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-07T13:20:45.494765Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:20:45.494765Z digest=sha256:f7c24faa57d78c2661d4279eec46cbdac99d50e3a004d8be1c742fa796e29fe5

Observation 48240a19-5a1f-4a6d-9ae4-a7e6e0a34b9d · outbound

This paper cites On the existence of minimizers in shallow residual relu neural network optimization landscapes.

PADAM: Parallel averaged Adam reduces the error for stochastic optimization in scientific machine learning On the existence of minimizers in shallow residual relu neural network optimization landscapes

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:20:56.058509Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T13:20:45.533294Z digest=sha256:fb82fa3a080bf040e8ddb8019070ae9384f7186d398a553536515d80efb7b718

Observation 87d7e869-f19e-427e-8c0a-7b832eba50bc · outbound

This paper cites Averaged Adam accelerates stochastic optimization in the training of deep neural network approximations for partial differential equation and optimal control problems.

PADAM: Parallel averaged Adam reduces the error for stochastic optimization in scientific machine learning Averaged Adam accelerates stochastic optimization in the training of deep neural network approximations for partial differential equation and optimal control problems

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-07T13:20:45.699370Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:20:45.699370Z digest=sha256:c62510c716657971a0a5a56f2333dd26645275134ed5b1442e212b9733820b40

Observation 080081b6-b3f0-4147-9f37-997a168118e0 · outbound

This paper cites Central limit theorems for stochastic gradient descent with averaging for stable manifolds.

PADAM: Parallel averaged Adam reduces the error for stochastic optimization in scientific machine learning Central limit theorems for stochastic gradient descent with averaging for stable manifolds

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:20:55.827850Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T13:20:45.799000Z digest=sha256:3cdffb741406d08fbd325c5690c96f5be5658b41caeed8aeb68f0188ccc67011

Observation e29d8abd-3823-40d6-97db-4159f880da90 · outbound

This paper cites On the existence of optimal shallow feedforward networks with ReLU activation.

PADAM: Parallel averaged Adam reduces the error for stochastic optimization in scientific machine learning On the existence of optimal shallow feedforward networks with ReLU activation

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:20:55.531121Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T13:20:45.895036Z digest=sha256:8dfddc6ba237df998569932f6ec8e7bb1fd50c63cbc5f8f5df87bf58a4c7f78f

Observation a998fdc4-adee-4388-aae0-cab3df4962ea · outbound

This paper cites General multilevel adaptations for stochastic approximation algorithms of Robbins-Monro and Polyak-Ruppert type.Numer.

PADAM: Parallel averaged Adam reduces the error for stochastic optimization in scientific machine learning General multilevel adaptations for stochastic approximation algorithms of Robbins-Monro and Polyak-Ruppert type.Numer

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:20:55.228206Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T13:20:46.038818Z digest=sha256:33cf2eae81b0e8a5612fef453bbe746d77e9d041a1fa64c010f1ef14a531696e

Observation 7765285a-e5f8-44e7-8862-38d86e513211 · outbound

This paper cites Uniform convergence guarantees for the deep ritz method for nonlinear problems.

PADAM: Parallel averaged Adam reduces the error for stochastic optimization in scientific machine learning Uniform convergence guarantees for the deep ritz method for nonlinear problems

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:20:54.928881Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T13:20:46.144661Z digest=sha256:2c93db44f07f1234030e0799b45ac900a2d9e7f04d0d96cfa2f1c10fde31f1a8

Observation 2e25f9d3-d276-4245-8d69-89a207cbf049 · outbound

This paper cites Deep learning-based numerical methods for high- dimensional parabolic partial differential equations and backward stochastic differential equations.

PADAM: Parallel averaged Adam reduces the error for stochastic optimization in scientific machine learning Deep learning-based numerical methods for high- dimensional parabolic partial differential equations and backward stochastic differential equations

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:20:54.610718Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T13:20:46.247445Z digest=sha256:fc833b8a82f8a73e69e5bc1a06f50c9c7d877aea1ad3e770f874e9b700828aba

Observation dd51ade3-a143-42ee-869d-c433c46f5dce · outbound

This paper cites Algorithms for solving high dimensional PDEs: from nonlinear Monte Carlo to machine learning.

PADAM: Parallel averaged Adam reduces the error for stochastic optimization in scientific machine learning Algorithms for solving high dimensional PDEs: from nonlinear Monte Carlo to machine learning

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:20:54.327060Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T13:20:46.383740Z digest=sha256:fc9fb22e25962a7db6fe5e089f62927e49d86454a0977a0642c4320857a8f284

Observation e4bdcae6-5852-49e7-afb1-ee27a3958610 · outbound

This paper cites The deep Ritz method: a deep learning-based numerical algorithm for solving variational problems.

PADAM: Parallel averaged Adam reduces the error for stochastic optimization in scientific machine learning The deep Ritz method: a deep learning-based numerical algorithm for solving variational problems

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:20:54.005662Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T13:20:46.481643Z digest=sha256:0e5d156e6a9f27440af31af2d62e8bce59f45fd6f1d334cc009c55619c65c8dd

Observation 91ddcca1-97a5-4f04-9fca-01d5749814f3 · outbound

This paper cites Optimal non-asymptotic bound of the Ruppert-Polyak averaging without strong convexity.

PADAM: Parallel averaged Adam reduces the error for stochastic optimization in scientific machine learning Optimal non-asymptotic bound of the Ruppert-Polyak averaging without strong convexity

Reference 25

Resolution
verified exact
local_arxiv, observed 2026-08-07T13:20:51.292006Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T13:20:46.576139Z digest=sha256:6b2e0e7c408ce3514dec07ceddd8fe10f0e33dd8682e040132a87a3c7d7662ce

Observation 3df7c9c6-1abd-410a-bfde-b0a24629b2e4 · outbound

This paper cites Blow up phenomena for gradient descent optimization methods in the training of artificial neural networks.

PADAM: Parallel averaged Adam reduces the error for stochastic optimization in scientific machine learning Blow up phenomena for gradient descent optimization methods in the training of artificial neural networks

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-07T13:20:46.654207Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:20:46.654207Z digest=sha256:678eb17da1a8ccb06d57aca9e5d6b97a22b0e0f70efd0e33a0bfe77bbc70bc3e

Observation 40721e7e-24dd-4a0c-856e-320d5fdbda52 · outbound

This paper cites Neural networks-based algorithms for stochastic control and PDEs in finance.

PADAM: Parallel averaged Adam reduces the error for stochastic optimization in scientific machine learning Neural networks-based algorithms for stochastic control and PDEs in finance

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-07T13:20:46.707687Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:20:46.707687Z digest=sha256:5d0d6cd477f3a84137826a53725e1c662e7ce734bec822d4665dbcc8f018386b

Observation 911bee71-a7b2-49a9-87b7-2c8b01730968 · outbound

This paper cites Stochastic weight averaging revisited.

PADAM: Parallel averaged Adam reduces the error for stochastic optimization in scientific machine learning Stochastic weight averaging revisited

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:20:53.734446Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T13:20:46.832622Z digest=sha256:09cc325e6fb14acec81d3bf91848cfd44965e4200a7ebeca9e45ca32b40328db

Observation 74dad194-4c96-4a2f-844e-990a800d4b01 · outbound

This paper cites an unresolved cited work.

PADAM: Parallel averaged Adam reduces the error for stochastic optimization in scientific machine learning Unresolved cited work

Reference 29

Resolution
unresolved
raw_fallback, observed 2026-08-07T13:20:53.453359Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T13:20:46.957269Z digest=sha256:7e7c8987d3df7e8021d9974267e69934acc441ad06a174b5d28d12e93bf1c7f1

Observation dd6775d0-2b8e-4823-ac1a-76cdd95391a3 · outbound

This paper cites Recent developments in machine learning methods for stochastic control and games.

PADAM: Parallel averaged Adam reduces the error for stochastic optimization in scientific machine learning Recent developments in machine learning methods for stochastic control and games

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:20:53.089257Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T13:20:47.034083Z digest=sha256:620e2a765711782834824122ad8a7dc51d659afc49e2ad86d101a91883d8b898

Observation f24dec9c-4fdd-40dd-a4a8-135a1fe8edcd · outbound

This paper cites Averaging Weights Leads to Wider Optima and Better Generalization.

PADAM: Parallel averaged Adam reduces the error for stochastic optimization in scientific machine learning Averaging Weights Leads to Wider Optima and Better Generalization

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-07T13:20:47.187971Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:20:47.187971Z digest=sha256:4707b1a6909c53eb5a01b32095b4c4842e91abe5b1964f1f0e0aa4cf2e6e2ac0

Observation 8916bea2-64ea-48d1-b5c0-54fa55f3d09e · outbound

This paper cites Mathematical Introduction to Deep Learning: Methods, Implementations, and Theory.

PADAM: Parallel averaged Adam reduces the error for stochastic optimization in scientific machine learning Mathematical Introduction to Deep Learning: Methods, Implementations, and Theory

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-07T13:20:47.254649Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:20:47.254649Z digest=sha256:9e47bee4cfed950b1c551198815c456f7aad8549c843539088c47fd7f99cc79a

Observation e75e8d49-a8c5-49ac-91b3-de4e0cc451f8 · outbound

This paper cites On the existence of global minima and convergence analyses for gradient descent methods in the training of deep neural networks.

PADAM: Parallel averaged Adam reduces the error for stochastic optimization in scientific machine learning On the existence of global minima and convergence analyses for gradient descent methods in the training of deep neural networks

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:20:52.838309Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T13:20:47.384573Z digest=sha256:292a14aba2ffce41c1462e4c5b912b96301a862968f6988738ddcc01d494f8bd

Observation 71cdcc58-6054-416f-b9c1-e954c0495dd4 · outbound

This paper cites Adam: A Method for Stochastic Optimization.

PADAM: Parallel averaged Adam reduces the error for stochastic optimization in scientific machine learning Adam: A Method for Stochastic Optimization

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-07T13:20:47.424978Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:20:47.424978Z digest=sha256:49a466bbae0eabab644e5134523f893f3aedc2239aaa91ab1b45af581c1ba44a

Observation d50492e0-b898-4191-a8c6-b43068b2341c · outbound

This paper cites SAD Neural Net- works: Divergent Gradient Flows and Asymptotic Optimality via o-minimal Structures.

PADAM: Parallel averaged Adam reduces the error for stochastic optimization in scientific machine learning SAD Neural Net- works: Divergent Gradient Flows and Asymptotic Optimality via o-minimal Structures

Reference 35

Resolution
verified exact
raw_fallback, observed 2026-08-07T13:20:50.894625Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T13:20:47.515572Z digest=sha256:9dcbea93e34fc55ab2ffddccfa9937456fde0acbf75ff0ed04b97af253edb038

Observation 4b85d118-57fa-461d-80bd-5ab60eaed08c · outbound

This paper cites Convergence of Adam Under Relaxed Assumptions.

PADAM: Parallel averaged Adam reduces the error for stochastic optimization in scientific machine learning Convergence of Adam Under Relaxed Assumptions

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-07T13:20:47.571118Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:20:47.571118Z digest=sha256:93289d74e54965aa1279ecce8ccf924e39d01802662445f933acb13e4fb000b4

Observation 8102fa5d-3cb1-4def-8a85-973b59d35ec3 · outbound

This paper cites Understanding SGD with Exponential Moving Average: A Case Study in Linear Regression.

PADAM: Parallel averaged Adam reduces the error for stochastic optimization in scientific machine learning Understanding SGD with Exponential Moving Average: A Case Study in Linear Regression

Reference 37

Resolution
verified exact
local_arxiv, observed 2026-08-07T13:20:50.561668Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T13:20:47.643359Z digest=sha256:3a2de582d33576339e4a5dff604dfff9e56411d40d814f242352d40a818241bc

Observation e11fa63e-2b52-4fda-8849-89de09810771 · outbound

This paper cites Summary of ChatGPT-Related Research and Perspective Towards the Future of Large Language Models.

PADAM: Parallel averaged Adam reduces the error for stochastic optimization in scientific machine learning Summary of ChatGPT-Related Research and Perspective Towards the Future of Large Language Models

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-07T13:20:47.730894Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:20:47.730894Z digest=sha256:584dbeff80e6cbf23fc92dd2509969c93510f1da8a5579eddc2bb78106a1916f

Observation dd8c1174-40ba-4975-a786-50ceab7bdbc9 · outbound

This paper cites Decoupled Weight Decay Regularization.

PADAM: Parallel averaged Adam reduces the error for stochastic optimization in scientific machine learning Decoupled Weight Decay Regularization

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-07T13:20:47.787142Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:20:47.787142Z digest=sha256:3174d967b309445fa95dbdb9f77ae5c774bb26ba5f89a94daccaabca36e599c6

Observation ff0b6f8b-69f6-458d-8324-e607f2d38ff7 · outbound

This paper cites Gradient Descent Maximizes the Margin of Homogeneous Neural Networks.

PADAM: Parallel averaged Adam reduces the error for stochastic optimization in scientific machine learning Gradient Descent Maximizes the Margin of Homogeneous Neural Networks

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-07T13:20:47.870928Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:20:47.870928Z digest=sha256:87b95226f70f0df8b80f418dec2ff65233b0e741d1dbd5ac37617b10cdaadae5

Observation f5a2318d-ef61-414c-80ee-bb1a7df5a500 · outbound

This paper cites Stochastic Gradient Descent as Approximate Bayesian Inference.

PADAM: Parallel averaged Adam reduces the error for stochastic optimization in scientific machine learning Stochastic Gradient Descent as Approximate Bayesian Inference

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-07T13:20:47.962359Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:20:47.962359Z digest=sha256:ef89bb4c3bf31329494b6c51873afae337e7bc6a9cfada4e591ae6f183935488

Observation 5c447aaa-cef6-4293-a0f0-4205e3852af2 · outbound

This paper cites Exponential moving average of weights in deep learning: Dynamics and benefits.

PADAM: Parallel averaged Adam reduces the error for stochastic optimization in scientific machine learning Exponential moving average of weights in deep learning: Dynamics and benefits

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:20:52.605227Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T13:20:48.034105Z digest=sha256:7b6209cf32c70ce4bee3b84fe81cac96ca6efbb478196b76dccf101a6879522b

Observation 394582bf-e2c6-4e71-8955-e6c29c87a6dd · outbound

This paper cites Topological properties of the set of functions generated by neural networks of fixed size.

PADAM: Parallel averaged Adam reduces the error for stochastic optimization in scientific machine learning Topological properties of the set of functions generated by neural networks of fixed size

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:20:52.390186Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T13:20:48.183493Z digest=sha256:3387ab49678ce92773b478e54151f6d846c0039180243630e2589ab820719359

Observation 0d60ffcf-2a5e-498c-96fa-3b4ddff012c6 · outbound

This paper cites Continuous-time stochastic control and optimization with financial applications, vol.

PADAM: Parallel averaged Adam reduces the error for stochastic optimization in scientific machine learning Continuous-time stochastic control and optimization with financial applications, vol

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:20:52.189882Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T13:20:48.301435Z digest=sha256:543b0b224e504b410e085f867876101fd58b16abc26d36d623968ba1a73406c5

Observation 20f15452-6034-473a-ba11-f349d65d8dc8 · outbound

This paper cites an unresolved cited work.

PADAM: Parallel averaged Adam reduces the error for stochastic optimization in scientific machine learning Unresolved cited work

Reference 45

Resolution
unresolved
raw_fallback, observed 2026-08-07T13:20:52.012419Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T13:20:48.411360Z digest=sha256:70f8f96db04b1c429e98944194c6edc57b8907ce0650a02b938fbedf4c88666d

Observation dd5bac77-c224-42ca-91f4-df1fe0e02d0c · outbound

This paper cites T., and Juditsky, A.

PADAM: Parallel averaged Adam reduces the error for stochastic optimization in scientific machine learning T., and Juditsky, A

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:20:51.839610Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T13:20:48.626059Z digest=sha256:638c6f880fc7a6babed54d446873f82e79fa3e9e1917779d569c18e2ec26dbaf

Observation 821a4347-7511-40bb-aac1-0a49b15ef796 · outbound

This paper cites Zero-Shot Text-to-Image Generation.

PADAM: Parallel averaged Adam reduces the error for stochastic optimization in scientific machine learning Zero-Shot Text-to-Image Generation

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-07T13:20:48.749797Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:20:48.749797Z digest=sha256:431f7f05ee4230a26a89cc96914382af5f6691303112f94af4b4f1103214c46a

Observation 8de2a81c-251e-457e-ad7f-3f692c0d2f2d · outbound

This paper cites On the Convergence of Adam and Beyond.

PADAM: Parallel averaged Adam reduces the error for stochastic optimization in scientific machine learning On the Convergence of Adam and Beyond

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-07T13:20:48.882281Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:20:48.882281Z digest=sha256:a71cc1980d7cd84667bc31fa76bdbb6d0a0fe8a7b18a1c7e67f4298f637e5dd3

Observation 5ceea0a7-d7e6-42e1-83a8-f50dc56cd332 · outbound

This paper cites High-Resolution Image Synthesis with Latent Diffusion Models.

PADAM: Parallel averaged Adam reduces the error for stochastic optimization in scientific machine learning High-Resolution Image Synthesis with Latent Diffusion Models

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-07T13:20:49.032440Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:20:49.032440Z digest=sha256:aadc802b1b9290c7f85eabcf1edb265dd213f3086a92b9e314d6d1e15b77c645

Observation 08ca1edf-a5b2-41d6-998e-8c2c25e34313 · outbound

This paper cites An overview of gradient descent optimization algorithms.

PADAM: Parallel averaged Adam reduces the error for stochastic optimization in scientific machine learning An overview of gradient descent optimization algorithms

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-07T13:20:49.150415Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:20:49.150415Z digest=sha256:e49406e2d4bc1fde357164898f28d7692c2a50a62e4e266dd2490e5dca6e2684

Observation b7389969-55b8-4b52-9f5c-8df99e018d28 · outbound

This paper cites Efficient estimations from a slowly convergent Robbins-Monro process.Cor- nell University Operations Research and Industrial Engineering, hdl.handle.net/1813/8664 (1988), 1–34.

PADAM: Parallel averaged Adam reduces the error for stochastic optimization in scientific machine learning Efficient estimations from a slowly convergent Robbins-Monro process.Cor- nell University Operations Research and Industrial Engineering, hdl.handle.net/1813/8664 (1988), 1–34

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:20:51.601182Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T13:20:49.266841Z digest=sha256:0ddbd7419fe1826ce401a888d0108b1273587061cb1c855a46c0bc8556618fb7

Observation 9ae9a0e3-f7c2-4cd3-a529-38c93cbaaf54 · outbound

This paper cites Photorealistic Text-to-Image Diffusion Models with Deep Language Understanding.

PADAM: Parallel averaged Adam reduces the error for stochastic optimization in scientific machine learning Photorealistic Text-to-Image Diffusion Models with Deep Language Understanding

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-07T13:20:49.372617Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:20:49.372617Z digest=sha256:38ed576d5ecd8704adae8a72e57d069246e06919d6a3e4c090e2854491128020

Observation 59a99ae4-2b76-4bb0-ad7f-ecea03b0c639 · outbound

This paper cites Training trajectories, mini-batch losses and the curious role of the learning rate.

PADAM: Parallel averaged Adam reduces the error for stochastic optimization in scientific machine learning Training trajectories, mini-batch losses and the curious role of the learning rate

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-07T13:20:49.539468Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:20:49.539468Z digest=sha256:231b3c218559b96593e2993b569eb35e42d2ed12247101b102642c621ec8bb36

Observation 503a1d71-526c-4b3a-8797-e03a34ddf718 · outbound

This paper cites On Margin Maximization in Linear and ReLU Networks.

PADAM: Parallel averaged Adam reduces the error for stochastic optimization in scientific machine learning On Margin Maximization in Linear and ReLU Networks

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-07T13:20:49.699637Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:20:49.699637Z digest=sha256:d6b957910bb37ffca4404deada7c09750d6c09bad6d4525ffffc01610d43f038

Observation de806598-2911-43d2-bd22-b44927348aae · outbound

This paper cites Deep learning with Elastic Averaging SGD.

PADAM: Parallel averaged Adam reduces the error for stochastic optimization in scientific machine learning Deep learning with Elastic Averaging SGD

Reference 55

Resolution
verified exact
local_arxiv, observed 2026-08-07T13:20:50.167971Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T13:20:49.832951Z digest=sha256:f5f0c2a9190eaf23611129ed8a159d7475af281f86744441ab9fefe577762418

Pith citing papers

Observation c4250ee0-78b1-45b3-8508-61ed43f75d3c · inbound

On the Provable Suboptimality of Momentum SGD in Nonstationary Stochastic Optimization cites this paper.

On the Provable Suboptimality of Momentum SGD in Nonstationary Stochastic Optimization PADAM: Parallel averaged Adam reduces the error for stochastic optimization in scientific machine learning

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-03T10:09:36.912553Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T10:09:36.912553Z digest=sha256:e6823a209b7f1c7e819a3f0af0841672133fd586f07977de49c383588514afcd

Observation b0941b4b-8e5c-49ff-ba18-6a59519f5548 · inbound

Central limit theorem for the averaged Adam optimizer cites this paper.

Central limit theorem for the averaged Adam optimizer PADAM: Parallel averaged Adam reduces the error for stochastic optimization in scientific machine learning

Reference 20

Resolution
verified exact
arxiv_id, observed 2026-07-04T07:29:38.374677Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-26T13:29:08.750421Z digest=sha256:5ff17c594c4572458043aa9bb3f2d79ee9d4b4bcbcf024525d79f6478bad16ac