Pith. sign in

Paper Citation Record · LEDGER

AutoSGD: Automatic Learning Rate Selection for Stochastic Gradient Descent

As of 18 August 2026, this Paper Citation Record lists 56 of 56 outbound references and 0 inbound Pith citation observations for arXiv:2505.21651.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.21651 v1

Coverage vector

measured 56 of 56 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T13:30:29.286244Z

measured 56 of 56 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-18T06:34:40.430872+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

56 of 56 outbound references displayed

  • verified exact2
  • verified fuzzy42
  • unresolved12
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 4ddf0f5f-4ee3-4ef8-8117-8dfa3e70f1be · outbound

This paper cites How Free is Parameter-Free Stochastic Optimization?.

AutoSGD: Automatic Learning Rate Selection for Stochastic Gradient Descent How Free is Parameter-Free Stochastic Optimization?

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-07T13:30:23.064115Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:30:23.064115Z digest=sha256:414e38ccf02c512dd80d04b1ab9c0d8c382df8cac6f078191a5b76b630fb8938

Observation aee387a0-0d2a-4cc7-9874-868a878da780 · outbound

This paper cites Calculus , volume 1.

AutoSGD: Automatic Learning Rate Selection for Stochastic Gradient Descent Calculus , volume 1

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:30:38.322528Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-07T13:30:23.187262Z digest=sha256:c43f838003c1131747864769722d30416cf2c37d8b30d2b6c262d17e3b6c2ca4

Observation fd78edb9-5eaa-446c-be1a-9cf9327857e3 · outbound

This paper cites Gradient descent converges linearly for logistic regression on separable data.

AutoSGD: Automatic Learning Rate Selection for Stochastic Gradient Descent Gradient descent converges linearly for logistic regression on separable data

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:30:38.105432Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-07T13:30:23.284766Z digest=sha256:ceaf6bd4cb0e91b5f206c2d106eb656c4410b51fc5e255a7b8e76d927f33d613

Observation 204a1cdd-8cea-4e8b-a5b1-2ac44fe571ee · outbound

This paper cites Julia: A fresh approach to numerical computing.

AutoSGD: Automatic Learning Rate Selection for Stochastic Gradient Descent Julia: A fresh approach to numerical computing

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:30:37.889146Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-07T13:30:23.335801Z digest=sha256:40779de4975cc0e384b07e08e7f8bf16afa376a79b5ea64c68fad4e8b69d849a

Observation dc942eac-f20d-4881-9c73-6ccf8118e61b · outbound

This paper cites Making SGD parameter-free.

AutoSGD: Automatic Learning Rate Selection for Stochastic Gradient Descent Making SGD parameter-free

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:30:37.708704Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-07T13:30:23.434935Z digest=sha256:4db75cf0a38b8e2b12dd0af09a53a10552ee9f4488324df8cd4a5bdb54848f46

Observation 649b4055-4bd7-4127-91ae-22caed5799c1 · outbound

This paper cites Understanding and detecting convergence for stochastic gradient descent with momentum.

AutoSGD: Automatic Learning Rate Selection for Stochastic Gradient Descent Understanding and detecting convergence for stochastic gradient descent with momentum

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:30:37.455861Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-07T13:30:23.514854Z digest=sha256:d3bf353b17dd475c623478520609f2ef964268367959dc42703fc9f727048845

Observation bfe847ce-4044-43d8-bb02-0ea0237f626a · outbound

This paper cites Convergence diagnostics for stochastic gradient descent with constant step size.

AutoSGD: Automatic Learning Rate Selection for Stochastic Gradient Descent Convergence diagnostics for stochastic gradient descent with constant step size

Reference 7

Resolution
verified exact
local_arxiv, observed 2026-08-07T13:30:29.874881Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-07T13:30:23.643611Z digest=sha256:85468fe36342a30b7b86b64e4d400ced99a796c288eb44c524eba0714c2f3642

Observation 42083acd-ec66-4610-9a39-1d291963493f · outbound

This paper cites Automatically constructing a corpus of sentential paraphrases.

AutoSGD: Automatic Learning Rate Selection for Stochastic Gradient Descent Automatically constructing a corpus of sentential paraphrases

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:30:37.288214Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-07T13:30:23.724831Z digest=sha256:66389fee84c630c002092d3fbc72b6b4e42921af1b7232b4dff62aff59272d16

Observation 9a2ecdec-5fd7-49f3-80ca-fc787084035f · outbound

This paper cites Robust, accurate stochastic optimization for variational inference.

AutoSGD: Automatic Learning Rate Selection for Stochastic Gradient Descent Robust, accurate stochastic optimization for variational inference

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:30:36.998998Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-07T13:30:23.809334Z digest=sha256:31437468a3ea1fb926c1ef2ba52b867d54ffd9f166ad03ae9c435165a93777b0

Observation 4171f955-c560-453a-a9de-dfcea1481669 · outbound

This paper cites Adaptive subgradient methods for online learning and stochastic optimization.

AutoSGD: Automatic Learning Rate Selection for Stochastic Gradient Descent Adaptive subgradient methods for online learning and stochastic optimization

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-07T13:30:23.876555Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:30:23.876555Z digest=sha256:4f86d5b09c461969d78efb9fcd9d8deea73f5d020a74761a17923ecbb4eac1cc

Observation 574adbda-bd38-4bb3-94fa-ad42f88a16f6 · outbound

This paper cites Learning-rate-free learning by D - A daptation.

AutoSGD: Automatic Learning Rate Selection for Stochastic Gradient Descent Learning-rate-free learning by D - A daptation

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:30:36.763690Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-07T13:30:23.965152Z digest=sha256:5e82fb1665dadc8aa02a32758f6dde8365fe011887600fb9fcdc7dad5015e4ca

Observation af7dafda-77e2-414a-8acb-756552a0c142 · outbound

This paper cites Markov Chains.

AutoSGD: Automatic Learning Rate Selection for Stochastic Gradient Descent Markov Chains

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:30:36.610874Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-07T13:30:24.009501Z digest=sha256:bbfb254f5163ed8ca83a8ab190c1f2721e5397d79cbd0142a9158adb828850b3

Observation aa54e8af-cc3b-455b-9b0e-3665668e6992 · outbound

This paper cites Probability: Theory and Examples.

AutoSGD: Automatic Learning Rate Selection for Stochastic Gradient Descent Probability: Theory and Examples

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:30:36.444990Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-07T13:30:24.084769Z digest=sha256:afec124309a32675ba3000c9bcceea9787141a3adbdc6c53f6f3bd988d3c410f

Observation b622ca89-67b0-408d-b08e-385de1e777df · outbound

This paper cites The road less scheduled.

AutoSGD: Automatic Learning Rate Selection for Stochastic Gradient Descent The road less scheduled

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:30:36.211598Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-07T13:30:24.161205Z digest=sha256:1e1c685473194f5546150c0216cc39a80bd41aa85743840c095282be6b7ad02e

Observation abd233f5-8329-4f89-8b88-1d11de747c4e · outbound

This paper cites Bayesian Data Analysis.

AutoSGD: Automatic Learning Rate Selection for Stochastic Gradient Descent Bayesian Data Analysis

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:30:36.005226Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-07T13:30:24.217374Z digest=sha256:6c824286f01b2dc4146c19d70e35b1f8f40e3949e79b68eb178847a03353a46a

Observation 7d99c6f7-0b53-42b8-a1aa-b86f26760ca8 · outbound

This paper cites Handbook of Convergence Theorems for (Stochastic) Gradient Methods.

AutoSGD: Automatic Learning Rate Selection for Stochastic Gradient Descent Handbook of Convergence Theorems for (Stochastic) Gradient Methods

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-07T13:30:24.271122Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:30:24.271122Z digest=sha256:1e2b31a1a0cffc9647842d2e4ed1c69a3e8919e232e683b2fddabfdf252d0088

Observation f3906197-8710-4483-afbb-550483e6c94e · outbound

This paper cites Inference from iterative simulation using multiple sequences.

AutoSGD: Automatic Learning Rate Selection for Stochastic Gradient Descent Inference from iterative simulation using multiple sequences

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:30:35.833611Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-07T13:30:24.344053Z digest=sha256:2846ac771c3c18d3e55963e4016e72f519c67a6b97fda95ed1e213ccdf3fb03c

Observation f269af48-9b48-43d1-983c-85bc5d2e6d37 · outbound

This paper cites Don't be so monotone: R elaxing stochastic line search in over-parameterized models.

AutoSGD: Automatic Learning Rate Selection for Stochastic Gradient Descent Don't be so monotone: R elaxing stochastic line search in over-parameterized models

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:30:35.705094Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-07T13:30:24.514784Z digest=sha256:5bf5a5dcebd3746314a9fb9f75df653aa2c69c44bf7990ac92bc26aab610f127

Observation 0380042f-3f8f-460a-a8b9-8bc45ae1a90f · outbound

This paper cites Variance-reduced methods for machine learning.

AutoSGD: Automatic Learning Rate Selection for Stochastic Gradient Descent Variance-reduced methods for machine learning

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:30:35.526086Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-07T13:30:24.644887Z digest=sha256:a1b000af5e3d00e5232c6f3208f2829709b762c7ff278931b6f955c80fd69ef8

Observation 36e6b8fe-9663-4636-b810-db23e70870ae · outbound

This paper cites Srivastava, and K.

AutoSGD: Automatic Learning Rate Selection for Stochastic Gradient Descent Srivastava, and K

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:30:35.292698Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-07T13:30:24.734965Z digest=sha256:e3903be744b481eaa6e56d45d9d73443c5c0c5345880568778fb1162a5f67a22

Observation fb0b79d0-6103-4375-8ba0-994978c72d94 · outbound

This paper cites Deep residual learning for image recognition.

AutoSGD: Automatic Learning Rate Selection for Stochastic Gradient Descent Deep residual learning for image recognition

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-07T13:30:24.804785Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:30:24.804785Z digest=sha256:f95a065b9a8f1b7c10896dad5303ac90038ae47de82632a7e3fa798f8989670b

Observation e243249d-3c26-45c5-80ad-d1392bbc1618 · outbound

This paper cites DoG is SGD 's best friend: A parameter-free dynamic step size schedule.

AutoSGD: Automatic Learning Rate Selection for Stochastic Gradient Descent DoG is SGD 's best friend: A parameter-free dynamic step size schedule

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:30:35.106516Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-07T13:30:24.917999Z digest=sha256:0f6af7317d98e564a171b0b2fbca0b7061cbd8a18b593caeba0fb5f7228367c0

Observation a7c3fab8-dca2-42b3-b1d3-4e9e0e683339 · outbound

This paper cites Parallelizing stochastic gradient descent for least squares regression: mini-batching, averaging, and model misspecification.

AutoSGD: Automatic Learning Rate Selection for Stochastic Gradient Descent Parallelizing stochastic gradient descent for least squares regression: mini-batching, averaging, and model misspecification

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:30:34.900072Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-07T13:30:25.014606Z digest=sha256:baad342aa05834aa2c087d98cdf4cd1098b1d29dd71ccf83a24fa65dfa6d500e

Observation 8c143f95-dbd4-4722-969e-b57ec402f1a1 · outbound

This paper cites Adam: A Method for Stochastic Optimization.

AutoSGD: Automatic Learning Rate Selection for Stochastic Gradient Descent Adam: A Method for Stochastic Optimization

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-07T13:30:25.236602Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:30:25.236602Z digest=sha256:bf238247c572c8d96b67e76f7592774da2c155f6d9fbefaf0db04612d143ea9b

Observation e9f5da1e-7290-4058-a06b-8bc3a8160200 · outbound

This paper cites Accelerated parameter-free stochastic optimization.

AutoSGD: Automatic Learning Rate Selection for Stochastic Gradient Descent Accelerated parameter-free stochastic optimization

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:30:34.699088Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-07T13:30:25.364739Z digest=sha256:e2f3caa130ee86071081e227a5b5b6db154ee9422573b64572be3f913e788831

Observation f1dbf6ab-fc57-427a-9682-0f28089a6926 · outbound

This paper cites Tuning-Free Stochastic Optimization.

AutoSGD: Automatic Learning Rate Selection for Stochastic Gradient Descent Tuning-Free Stochastic Optimization

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-07T13:30:25.644802Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:30:25.644802Z digest=sha256:db494c6b1b8ac84a2da34bb573d21d9c189748f10bfb397db8197e7efe1b3021

Observation 332916ce-44bc-4a4c-a515-c14d1a4eabfb · outbound

This paper cites Linear convergence of black-box variational inference: S hould we stick the landing? In International Conference on Artificial Intelligence and Statistics , pages 235--243.

AutoSGD: Automatic Learning Rate Selection for Stochastic Gradient Descent Linear convergence of black-box variational inference: S hould we stick the landing? In International Conference on Artificial Intelligence and Statistics , pages 235--243

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:30:34.500231Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-07T13:30:25.815236Z digest=sha256:a8433fccfa7a1cda706dc4bcee3b57a0572c1a1a3a19924b64069f3de9d3adbb

Observation cc97eaae-22fe-433a-a526-03e65b9c99e0 · outbound

This paper cites DoWG unleashed: A n efficient universal parameter-free gradient descent method.

AutoSGD: Automatic Learning Rate Selection for Stochastic Gradient Descent DoWG unleashed: A n efficient universal parameter-free gradient descent method

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:30:34.345067Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-07T13:30:25.935152Z digest=sha256:b8593cc15846987a7cd4627a2c041d08500249cfd99f08ad2801b51556bafeab

Observation 3d872e16-63cf-40c7-8062-5de2ee988c80 · outbound

This paper cites Learning multiple layers of features from tiny images.

AutoSGD: Automatic Learning Rate Selection for Stochastic Gradient Descent Learning multiple layers of features from tiny images

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-07T13:30:26.009608Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:30:26.009608Z digest=sha256:745e371f3ab7739a58b90f184c9a93fc864a4f4db6ec4eeb8406ce99f5c73502

Observation 3a8ed22e-5ebb-4893-875d-715a7d6c99ab · outbound

This paper cites RoBERTa: A Robustly Optimized BERT Pretraining Approach.

AutoSGD: Automatic Learning Rate Selection for Stochastic Gradient Descent RoBERTa: A Robustly Optimized BERT Pretraining Approach

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-07T13:30:26.125015Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:30:26.125015Z digest=sha256:85389d2a2300162a271999f5881ab4612afb8fd44751f1358c39ed46aa5bac56

Observation 8666a607-6920-4c21-8721-f33055dabb10 · outbound

This paper cites Stochastic polyak step-size for SGD : A n adaptive learning rate for fast convergence.

AutoSGD: Automatic Learning Rate Selection for Stochastic Gradient Descent Stochastic polyak step-size for SGD : A n adaptive learning rate for fast convergence

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:30:34.051792Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-07T13:30:26.255069Z digest=sha256:3c23983efbe97fc133bfd0f9b9e3ba6578a4c597bd2d390608992687cb056ad1

Observation 6bb0fb86-b258-49aa-aa2b-d30b47fdaf7e · outbound

This paper cites Using statistics to automate stochastic optimization.

AutoSGD: Automatic Learning Rate Selection for Stochastic Gradient Descent Using statistics to automate stochastic optimization

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:30:33.896297Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-07T13:30:26.386965Z digest=sha256:472fd0227321df2fb5ecf0bd5796ea13fb514c61c914c5661363178ad1bba0b8

Observation ea3a5ea4-6857-4233-9231-459089494ea2 · outbound

This paper cites Prodigy: An Expeditiously Adaptive Parameter-Free Learner.

AutoSGD: Automatic Learning Rate Selection for Stochastic Gradient Descent Prodigy: An Expeditiously Adaptive Parameter-Free Learner

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-07T13:30:26.479605Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:30:26.479605Z digest=sha256:b6780469dbe10aa926769f3afd4a4be2bffef62e116e94f687c0322a33322ea6

Observation 62a50b7a-aa01-4eb5-a9bc-605cdc44fc4f · outbound

This paper cites Adaptive Gradient Descent without Descent.

AutoSGD: Automatic Learning Rate Selection for Stochastic Gradient Descent Adaptive Gradient Descent without Descent

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-07T13:30:26.665057Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:30:26.665057Z digest=sha256:9307265342a7834cd2649758b6033c86c87b3661c82abf79f1cd49e04d4ada20

Observation 23103646-e9af-4aaf-8a31-f8b19b14b154 · outbound

This paper cites Beating SGD saturation with tail-averaging and minibatching.

AutoSGD: Automatic Learning Rate Selection for Stochastic Gradient Descent Beating SGD saturation with tail-averaging and minibatching

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:30:33.733177Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-07T13:30:26.964874Z digest=sha256:cbbb1b44288dcc0e96d9aa1eea35c21b0f1c062c59df5ab5b5b10ec7e4010dad

Observation 662b3f75-7b27-4b96-8588-01f6aa9ba61f · outbound

This paper cites Let's make block coordinate descent converge faster: F aster greedy rules, message-passing, active-set complexity, and superlinear convergence.

AutoSGD: Automatic Learning Rate Selection for Stochastic Gradient Descent Let's make block coordinate descent converge faster: F aster greedy rules, message-passing, active-set complexity, and superlinear convergence

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:30:33.571767Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-07T13:30:27.065084Z digest=sha256:4e7f22850a11b96d10ce543daeaffab225314e091da362c608a59e5ce44e0e4e

Observation fcb7d2fb-32de-4f7e-a141-c063f88ba817 · outbound

This paper cites Dynamics of SGD with stochastic P olyak stepsizes: T ruly adaptive variants and convergence to exact solution.

AutoSGD: Automatic Learning Rate Selection for Stochastic Gradient Descent Dynamics of SGD with stochastic P olyak stepsizes: T ruly adaptive variants and convergence to exact solution

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:30:33.370986Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-07T13:30:27.182224Z digest=sha256:94dde148b14845ceddea66c2db9dff8f16371995c56fa6eaac52391e022e4ac4

Observation efce4528-360d-4278-8025-944fffac552d · outbound

This paper cites Training deep networks without learning rates through coin betting.

AutoSGD: Automatic Learning Rate Selection for Stochastic Gradient Descent Training deep networks without learning rates through coin betting

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:30:33.098648Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-07T13:30:27.414972Z digest=sha256:750d22e4ba63f44c85eef9500f37d07f5fefa7afd6416edece3ba05e7660e740

Observation b1f77e31-f3f6-4b17-aa45-471c6eeaebac · outbound

This paper cites On convergence-diagnostic based step sizes for stochastic gradient descent.

AutoSGD: Automatic Learning Rate Selection for Stochastic Gradient Descent On convergence-diagnostic based step sizes for stochastic gradient descent

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:30:32.885767Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-07T13:30:27.584877Z digest=sha256:fbd56a9cd7e4a8dc1e7fb3ef56c044da14a3b898b40ae91ed9214bf7b12e0892

Observation c25f0525-7fc9-440f-9ef4-81f3c216a59f · outbound

This paper cites On the determination of the step size in stochastic quasigradient methods.

AutoSGD: Automatic Learning Rate Selection for Stochastic Gradient Descent On the determination of the step size in stochastic quasigradient methods

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:30:32.639217Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-07T13:30:27.666340Z digest=sha256:cab32028538994a68fb37011181421e00195d2942dd3a5e6fa5eaed3e0811cf3

Observation b6ffb9cd-dac4-46c0-8cec-a7c171abf7f7 · outbound

This paper cites Non-asymptotic confidence bounds for stochastic approximation algorithms with constant step size.

AutoSGD: Automatic Learning Rate Selection for Stochastic Gradient Descent Non-asymptotic confidence bounds for stochastic approximation algorithms with constant step size

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:30:32.392834Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-07T13:30:27.774917Z digest=sha256:994f7bdaf4b0c455ae634d73e2651e9ad79937c0741718f55d093addf8d3886f

Observation 359e5e04-282e-4398-89dc-05087705d8ab · outbound

This paper cites Py T orch: A n imperative style, high-performance deep learning library.

AutoSGD: Automatic Learning Rate Selection for Stochastic Gradient Descent Py T orch: A n imperative style, high-performance deep learning library

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:30:32.241807Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-07T13:30:27.905036Z digest=sha256:d6c7fd7af50a99a44a7571035341189027930128000efeb274f44c82a3953ed6

Observation acf787f1-2a75-4fed-967d-d50e6342a6a0 · outbound

This paper cites https://huggingface.co/microsoft/resnet-18, 2025.

AutoSGD: Automatic Learning Rate Selection for Stochastic Gradient Descent https://huggingface.co/microsoft/resnet-18, 2025

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:30:32.018044Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-07T13:30:28.010649Z digest=sha256:299d7a7b8b1f95c05d4060534079b4402825d86aba9e2e554a84b6ff980b0d7c

Observation 6454f534-0c17-463f-9b38-3a1d916d3ddc · outbound

This paper cites A stochastic approximation method.

AutoSGD: Automatic Learning Rate Selection for Stochastic Gradient Descent A stochastic approximation method

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:30:31.783988Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-07T13:30:28.121239Z digest=sha256:a78fb5e28f3241fb8680efdcca2aa1c795ba2441b4d0b637b03beac76e7f7ac4

Observation 4a60d70a-43fc-4b69-9cd4-e010c9f1012e · outbound

This paper cites https://huggingface.co/FacebookAI/roberta-base, 2025.

AutoSGD: Automatic Learning Rate Selection for Stochastic Gradient Descent https://huggingface.co/FacebookAI/roberta-base, 2025

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:30:31.514873Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-07T13:30:28.463148Z digest=sha256:24570cbca3986e0d3d9135fd354d6dbbb8b04e4fc01c4e3f58e75b7c6da371d4

Observation 7927b180-2727-410e-b59e-a6b26519e75e · outbound

This paper cites Anytime Tail Averaging.

AutoSGD: Automatic Learning Rate Selection for Stochastic Gradient Descent Anytime Tail Averaging

Reference 46

Resolution
verified exact
local_arxiv, observed 2026-08-07T13:30:29.584753Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-07T13:30:28.540616Z digest=sha256:b2b2e4cccf969d453d25ef18c06e90d2058130a838e66027a955cbf6dbb2fda5

Observation dbfa0b42-1d36-4900-9839-ef3ce529f51a · outbound

This paper cites Making Gradient Descent Optimal for Strongly Convex Stochastic Optimization.

AutoSGD: Automatic Learning Rate Selection for Stochastic Gradient Descent Making Gradient Descent Optimal for Strongly Convex Stochastic Optimization

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-07T13:30:28.585346Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:30:28.585346Z digest=sha256:b4f03a9f5846bb6596c1d0ae93c99df325906c2c197e3fc0c34f3adef57f1228

Observation 2f10f48c-f16a-44cb-941d-d5d633ef795d · outbound

This paper cites Sticking the landing: S imple, lower-variance gradient estimators for variational inference.

AutoSGD: Automatic Learning Rate Selection for Stochastic Gradient Descent Sticking the landing: S imple, lower-variance gradient estimators for variational inference

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:30:31.347644Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-07T13:30:28.687107Z digest=sha256:5a09f419e8cbf7ec7f9be5f8fd4d3508acfa9821c5e87f0b667db4073086b1ec

Observation 14442b01-b4df-40e7-b73f-8a897ade788d · outbound

This paper cites SQuAD : 100,000+ questions for machine comprehension of text.

AutoSGD: Automatic Learning Rate Selection for Stochastic Gradient Descent SQuAD : 100,000+ questions for machine comprehension of text

Reference 49

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:30:31.130714Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-07T13:30:28.737065Z digest=sha256:142de3d15e5f7bc3fde4b22d9ac50f3c64dcdbe85dbcf5c649b641f52b84a6e5

Observation 8d42f786-343b-4ec1-8444-9576d3126183 · outbound

This paper cites Virtual library of simulation experiments: T est functions and datasets, 2013.

AutoSGD: Automatic Learning Rate Selection for Stochastic Gradient Descent Virtual library of simulation experiments: T est functions and datasets, 2013

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:30:30.933615Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-07T13:30:28.815036Z digest=sha256:8447cb7fe0f1a3969220f3e3da31e80465718bc00fc89e6b60befa126af1df21

Observation 0acc27dd-ad8a-41c7-8cbe-df18780bf2d4 · outbound

This paper cites Recursive deep models for semantic compositionality over a sentiment treebank.

AutoSGD: Automatic Learning Rate Selection for Stochastic Gradient Descent Recursive deep models for semantic compositionality over a sentiment treebank

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:30:30.762131Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-07T13:30:28.856405Z digest=sha256:6e249441da48f2d833271cf2d6dac54b5342fd1cbbd6f23938b3750ce86d6e1f

Observation b597a55d-446a-46dd-83f6-c4175b93c23b · outbound

This paper cites Stochastic gradient descent for non-smooth optimization: C onvergence results and optimal averaging schemes.

AutoSGD: Automatic Learning Rate Selection for Stochastic Gradient Descent Stochastic gradient descent for non-smooth optimization: C onvergence results and optimal averaging schemes

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:30:30.583910Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-07T13:30:28.911294Z digest=sha256:399b80dba0e193f6125a875dae5baf7712823e56fe668fb68bec0d10c3ab2626

Observation f81a50cd-0413-41b8-9408-88f6d4dd081f · outbound

This paper cites Painless stochastic gradient: I nterpolation, line-search, and convergence rates.

AutoSGD: Automatic Learning Rate Selection for Stochastic Gradient Descent Painless stochastic gradient: I nterpolation, line-search, and convergence rates

Reference 53

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:30:30.444768Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-07T13:30:28.941336Z digest=sha256:ba2658a8d1775d60b36251cbcc6dc964460ca8315d485c0fdaf4b9f13a96ea93

Observation 19139d0b-134b-46c5-a154-68bf21d15add · outbound

This paper cites A framework for improving the reliability of black-box variational inference.

AutoSGD: Automatic Learning Rate Selection for Stochastic Gradient Descent A framework for improving the reliability of black-box variational inference

Reference 54

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:30:30.280853Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-07T13:30:29.073747Z digest=sha256:2787044f4f3cc77db1fb23a81f110556c187a85f78b0a212a74df1f7f363a23d

Observation f258329d-7b20-498c-a848-44bc1f716d5b · outbound

This paper cites GLUE : A multi-task benchmark and analysis platform for natural language understanding.

AutoSGD: Automatic Learning Rate Selection for Stochastic Gradient Descent GLUE : A multi-task benchmark and analysis platform for natural language understanding

Reference 55

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:30:30.093008Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-07T13:30:29.175993Z digest=sha256:ab8a32482a991f13ed7847d37176022e41bf533286ad8e1879d7251840d3bd78

Observation 563ed8bc-d756-4a3a-a11e-8f9e3f48bb71 · outbound

This paper cites Fluctuation-dissipation relations for stochastic gradient descent.

AutoSGD: Automatic Learning Rate Selection for Stochastic Gradient Descent Fluctuation-dissipation relations for stochastic gradient descent

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-07T13:30:29.286244Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:30:29.286244Z digest=sha256:0b225f828d8fab7006e7f985b92254b223a485fd6382ac64b24e0e59bb394c68

Pith citing papers

No inbound Pith citation observations are available.