Pith. sign in

Paper Citation Record · LEDGER

LionVote: Per-Layer Learning Rate Adaptation for Lion

As of 12 August 2026, this Paper Citation Record lists 25 of 25 outbound references and 0 inbound Pith citation observations for arXiv:2607.09266.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2607.09266 v1

Coverage vector

measured 25 of 25 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-07-13T04:17:58.962415Z

measured 25 of 25 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-11T06:34:44.6726+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

25 of 25 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved25
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 8cbeb75d-54bd-407e-b200-5cb587aa8cc2 · outbound

This paper cites W., Pfau, D., Schaul, T., Shillingford, B., and de Freitas, N.

LionVote: Per-Layer Learning Rate Adaptation for Lion W., Pfau, D., Schaul, T., Shillingford, B., and de Freitas, N

Reference 1

Resolution
unresolved
no resolver link, observed 2026-07-13T04:17:58.962415Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-13T04:17:58.962415Z digest=sha256:de328b2b25fe4aba3b6efe0212442e20f6a94002dda62b4f0b8695903f41dfe0

Observation a25c34a1-1b4b-42a0-934a-2484819e7bfd · outbound

This paper cites signSGD : Compressed optimisation for non-convex problems.

LionVote: Per-Layer Learning Rate Adaptation for Lion signSGD : Compressed optimisation for non-convex problems

Reference 2

Resolution
unresolved
no resolver link, observed 2026-07-13T04:17:58.962415Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-13T04:17:58.962415Z digest=sha256:09a50dab6916cf6ba1dc807c276bd9ba0667d9e39d5c857795ac45db56f5b76d

Observation 5bdcdcfa-fd1c-4836-9257-62c448bfac8e · outbound

This paper cites an unresolved cited work.

LionVote: Per-Layer Learning Rate Adaptation for Lion Unresolved cited work

Reference 3

Resolution
unresolved
no resolver link, observed 2026-07-13T04:17:58.962415Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-13T04:17:58.962415Z digest=sha256:db99ed32fba95bd18f0239f63d83d7bc054fb82a524008300afb4c0f6946cfd2

Observation f6d98635-3bc1-4f66-bd4a-12b8920293a5 · outbound

This paper cites Z., and Talwalkar, A.

LionVote: Per-Layer Learning Rate Adaptation for Lion Z., and Talwalkar, A

Reference 4

Resolution
unresolved
no resolver link, observed 2026-07-13T04:17:58.962415Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-13T04:17:58.962415Z digest=sha256:9cb1c5df5e048539bcae986791e2b9c0f35903a0d13af2a6d80b1bce8cf11492

Observation 70787327-038a-456b-bae0-314679f04184 · outbound

This paper cites and Mishchenko, K.

LionVote: Per-Layer Learning Rate Adaptation for Lion and Mishchenko, K

Reference 5

Resolution
unresolved
no resolver link, observed 2026-07-13T04:17:58.962415Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-13T04:17:58.962415Z digest=sha256:344c1837540b8e449dbb55a660ebcde49df10c7aff8e05cbbabd3874af1cc613

Observation c3eef540-d207-4514-83a2-3e8c98749e23 · outbound

This paper cites The road less scheduled.

LionVote: Per-Layer Learning Rate Adaptation for Lion The road less scheduled

Reference 6

Resolution
unresolved
no resolver link, observed 2026-07-13T04:17:58.962415Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-13T04:17:58.962415Z digest=sha256:80b5b442f2fa076968b6f3da4578a44cca948a3cd92dbde8881fd7666de11fd2

Observation ba9c237d-58c4-43d8-a5b2-868dd2fb934b · outbound

This paper cites An image is worth 16x16 words: Transformers for image recognition at scale.

LionVote: Per-Layer Learning Rate Adaptation for Lion An image is worth 16x16 words: Transformers for image recognition at scale

Reference 7

Resolution
unresolved
no resolver link, observed 2026-07-13T04:17:58.962415Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-13T04:17:58.962415Z digest=sha256:d6b30ecf35fa0db42258ae93e125e9a2c108f6e72ba0d4ca1bfb0b0841b497de

Observation 08cf2aab-f2b3-4e81-9178-8056e98fc953 · outbound

This paper cites Adaptive subgradient methods for online learning and stochastic optimization.

LionVote: Per-Layer Learning Rate Adaptation for Lion Adaptive subgradient methods for online learning and stochastic optimization

Reference 8

Resolution
unresolved
no resolver link, observed 2026-07-13T04:17:58.962415Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-13T04:17:58.962415Z digest=sha256:efb4c424c7d096358e1bcb1dd00c6b7402ce00d713eb12282909d6bfc39e05b0

Observation c79e9384-e9e9-428c-ac9d-26edfcdc1110 · outbound

This paper cites Deep residual learning for image recognition.

LionVote: Per-Layer Learning Rate Adaptation for Lion Deep residual learning for image recognition

Reference 9

Resolution
unresolved
no resolver link, observed 2026-07-13T04:17:58.962415Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-13T04:17:58.962415Z digest=sha256:e3b9e3bbc81b8e040065f0809411e4a2a0fb4acfb1d1e17316d14d8dad467d17

Observation 4f3828e8-d530-4011-9151-3ab64144bf3d · outbound

This paper cites Noise-adaptive layerwise learning rates: Accelerating geometry-aware optimization for deep neural network training.

LionVote: Per-Layer Learning Rate Adaptation for Lion Noise-adaptive layerwise learning rates: Accelerating geometry-aware optimization for deep neural network training

Reference 10

Resolution
unresolved
no resolver link, observed 2026-07-13T04:17:58.962415Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-13T04:17:58.962415Z digest=sha256:5b0ee5c09f9e69df5ff4aa0fe945b500abb9fec9735329da5b8ce2a83a73cc56

Observation 4db50434-f295-4ccb-9b54-85a1307e4a6c · outbound

This paper cites and Ruder, S.

LionVote: Per-Layer Learning Rate Adaptation for Lion and Ruder, S

Reference 11

Resolution
unresolved
no resolver link, observed 2026-07-13T04:17:58.962415Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-13T04:17:58.962415Z digest=sha256:c26339c338205b19fba4accd3da7ff6b38b1d78643cd2bf8462a3f207b2fc962

Observation 06a976f0-f5bc-4e5a-83ef-8f5eb0e1c079 · outbound

This paper cites Muon: An optimizer for hidden layers in neural networks.

LionVote: Per-Layer Learning Rate Adaptation for Lion Muon: An optimizer for hidden layers in neural networks

Reference 12

Resolution
unresolved
no resolver link, observed 2026-07-13T04:17:58.962415Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-13T04:17:58.962415Z digest=sha256:cef0db16f9e6e3479bb292b5f6bac84ae6f611ea9408db863c8a5961ad5febc5

Observation 8110590f-a14b-43e2-b304-8267cd8fde8c · outbound

This paper cites an unresolved cited work.

LionVote: Per-Layer Learning Rate Adaptation for Lion Unresolved cited work

Reference 13

Resolution
unresolved
no resolver link, observed 2026-07-13T04:17:58.962415Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-13T04:17:58.962415Z digest=sha256:eeda588b1ec91c0aae79e1d9b1c2d578994e4396c6b025b8bd21b67363fc375c

Observation 6e190b7d-e15c-4c28-8802-d27bbc39398d · outbound

This paper cites Cautious optimizers: Improving training with one line of code.

LionVote: Per-Layer Learning Rate Adaptation for Lion Cautious optimizers: Improving training with one line of code

Reference 14

Resolution
unresolved
no resolver link, observed 2026-07-13T04:17:58.962415Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-13T04:17:58.962415Z digest=sha256:2254e609fac2b5ba49f7c155b4686efa5a00cccd33a58e98252cd1112a9d74dd

Observation 03d906d2-8255-4f6f-92c7-f7c5afb7113b · outbound

This paper cites and Hutter, F.

LionVote: Per-Layer Learning Rate Adaptation for Lion and Hutter, F

Reference 15

Resolution
unresolved
no resolver link, observed 2026-07-13T04:17:58.962415Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-13T04:17:58.962415Z digest=sha256:ef61dd453785a19a0446bcfa3a85b97dd2d3c7fc7d2eba7371641f04a6635e6a

Observation 933195f3-a949-4f29-abd7-513c4181c0f2 · outbound

This paper cites and Hutter, F.

LionVote: Per-Layer Learning Rate Adaptation for Lion and Hutter, F

Reference 16

Resolution
unresolved
no resolver link, observed 2026-07-13T04:17:58.962415Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-13T04:17:58.962415Z digest=sha256:c09883cdc33e6bec03a9925a23bf893da2b13fa3eeb1bcb84ca772cd78ec7b14

Observation 3fd542f2-6995-43d6-a295-1606c06ae251 · outbound

This paper cites PyTorch : An imperative style, high-performance deep learning library.

LionVote: Per-Layer Learning Rate Adaptation for Lion PyTorch : An imperative style, high-performance deep learning library

Reference 17

Resolution
unresolved
no resolver link, observed 2026-07-13T04:17:58.962415Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-13T04:17:58.962415Z digest=sha256:04493ab5948254645215de079fc053e5045bb4326631e411c6d8a20c234f0235

Observation 2f8aa75b-f52d-4c33-8878-41e9c90174c3 · outbound

This paper cites An Adaptive Stochastic Gradient Method with Non-negative Gauss-Newton Stepsizes.

LionVote: Per-Layer Learning Rate Adaptation for Lion An Adaptive Stochastic Gradient Method with Non-negative Gauss-Newton Stepsizes

Reference 18

Resolution
unresolved
no resolver link, observed 2026-07-13T04:17:58.962415Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-13T04:17:58.962415Z digest=sha256:79dfa12d99b3925f4206c4ba0b6ef0f3af2df2faa00934a8bfbf66cb36ec9e0f

Observation e959b445-3a50-4007-a07d-6cc5159f9fa0 · outbound

This paper cites M., Schneider, F., and Hennig, P.

LionVote: Per-Layer Learning Rate Adaptation for Lion M., Schneider, F., and Hennig, P

Reference 19

Resolution
unresolved
no resolver link, observed 2026-07-13T04:17:58.962415Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-13T04:17:58.962415Z digest=sha256:b7479df16bd75fa41fc39e6ed0aee09e10e3898a52386b32ccf058945a58aa65

Observation 2fb2fa15-5ce9-4da5-885e-850f11a83116 · outbound

This paper cites N., Kaiser, L., and Polosukhin, I.

LionVote: Per-Layer Learning Rate Adaptation for Lion N., Kaiser, L., and Polosukhin, I

Reference 20

Resolution
unresolved
no resolver link, observed 2026-07-13T04:17:58.962415Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-13T04:17:58.962415Z digest=sha256:e1e40964b685f663e25c1c60eca8ee6be96336cc2f9b6a6a70ec4fd863b10b01

Observation 08f26162-caec-4d28-932a-a891112e0e4c · outbound

This paper cites AutoDrop : Training deep learning models with automatic learning rate drop.

LionVote: Per-Layer Learning Rate Adaptation for Lion AutoDrop : Training deep learning models with automatic learning rate drop

Reference 21

Resolution
unresolved
no resolver link, observed 2026-07-13T04:17:58.962415Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-13T04:17:58.962415Z digest=sha256:b1b9e6b0750e8e2428e153b689eaab54561edc2d041f957af9ea49ce67f5319c

Observation 3753eee2-05cc-4415-96a8-1515526e9942 · outbound

This paper cites Large Batch Training of Convolutional Networks.

LionVote: Per-Layer Learning Rate Adaptation for Lion Large Batch Training of Convolutional Networks

Reference 22

Resolution
unresolved
no resolver link, observed 2026-07-13T04:17:58.962415Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-13T04:17:58.962415Z digest=sha256:b119317d4a8166f46c42f3f89914f21bf1a907536fd4758ebabfcdcd5b53ae06

Observation 8bff8c94-5546-4ab1-844a-34b1016ae3e4 · outbound

This paper cites Large batch optimization for deep learning: Training BERT in 76 minutes.

LionVote: Per-Layer Learning Rate Adaptation for Lion Large batch optimization for deep learning: Training BERT in 76 minutes

Reference 23

Resolution
unresolved
no resolver link, observed 2026-07-13T04:17:58.962415Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-13T04:17:58.962415Z digest=sha256:569c4bff3445fa87d6f4c013541796033e4c2538e72664e87fba0f07451343d3

Observation d75d1813-ca1e-414e-b112-a713f5e00568 · outbound

This paper cites and Komodakis, N.

LionVote: Per-Layer Learning Rate Adaptation for Lion and Komodakis, N

Reference 24

Resolution
unresolved
no resolver link, observed 2026-07-13T04:17:58.962415Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-13T04:17:58.962415Z digest=sha256:3e9447cc25e141c67e8c9120a1735c732d2c5898c2a2adf4645dfd4d92776bb8

Observation 28f9fdb4-e44b-4966-b24c-0d1186f6183a · outbound

This paper cites Deconstructing what makes a good optimizer for autoregressive language models.

LionVote: Per-Layer Learning Rate Adaptation for Lion Deconstructing what makes a good optimizer for autoregressive language models

Reference 25

Resolution
unresolved
no resolver link, observed 2026-07-13T04:17:58.962415Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-13T04:17:58.962415Z digest=sha256:70fbdcccb007ff9d3a2cec922dad94b2ab1558c817d9215b31e028f45935e743

Pith citing papers

No inbound Pith citation observations are available.