Pith. sign in

Paper Citation Record · LEDGER

Mathematical analysis of the gradients in deep learning

As of 11 August 2026, this Paper Citation Record lists 58 of 58 outbound references and 0 inbound Pith citation observations for arXiv:2501.15646.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2501.15646 v1

Coverage vector

measured 58 of 58 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-10T14:10:55.394691Z

measured 58 of 58 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

58 of 58 outbound references displayed

  • verified exact11
  • verified fuzzy22
  • unresolved24
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch1

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation c4197251-d85e-4647-b255-13d763e15d4b · outbound

This paper cites TensorFlow: Large-Scale Machine Learning on Heterogeneous Distributed Systems.

Mathematical analysis of the gradients in deep learning TensorFlow: Large-Scale Machine Learning on Heterogeneous Distributed Systems

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-10T14:10:55.075137Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T14:10:55.075137Z digest=sha256:acd9a29361c618fb2d3917b230e3ff7cf5ed88a36643f96386e3b79bfffc11cb

Observation 2cd1384d-7819-49a9-8206-f4fb72d35904 · outbound

This paper cites Learning Theory from First Principles.

Mathematical analysis of the gradients in deep learning Learning Theory from First Principles

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-10T14:10:55.079600Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T14:10:55.079600Z digest=sha256:5a09908b694d286a6aa316cb800739e702d861d32bcfc440cb5b367db10b014f

Observation 07e570e5-81b0-4bba-81a4-327baf331d8f · outbound

This paper cites J., and Zhang, Y.

Mathematical analysis of the gradients in deep learning J., and Zhang, Y

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T14:10:56.236616Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T14:10:55.083893Z digest=sha256:36ce8a804bf954170189e781a971e4a6725be3febd1cbcdf4887f7a26130b122

Observation 0bd24e4a-9259-4986-82f1-e79ec3902d01 · outbound

This paper cites On the complexity of nonsmooth automatic differentiation.

Mathematical analysis of the gradients in deep learning On the complexity of nonsmooth automatic differentiation

Reference 4

Resolution
verified exact
local_arxiv, observed 2026-08-10T14:10:55.897472Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T14:10:55.087847Z digest=sha256:0cf5283c1927e098751b8327a8ea0aa77ae527b4a53ae22f22453c05ee05edc0

Observation 386b18a8-919b-4bbd-98d3-137ca1f53ccf · outbound

This paper cites The /suppress lojasiewicz inequality for nonsmooth subanalytic functions with applications to subgradient dynamical systems.

Mathematical analysis of the gradients in deep learning The /suppress lojasiewicz inequality for nonsmooth subanalytic functions with applications to subgradient dynamical systems

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T14:10:56.224193Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T14:10:55.092333Z digest=sha256:909641169c198063fc53c206565189636c797bf9c4cd49de412d6f52bf5a8dc3

Observation 38756945-2ca7-4bfe-a275-52d5d8f00612 · outbound

This paper cites A mathematical model for automatic differentiation in machine learning.

Mathematical analysis of the gradients in deep learning A mathematical model for automatic differentiation in machine learning

Reference 6

Resolution
verified exact
local_arxiv, observed 2026-08-10T14:10:55.881349Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T14:10:55.096392Z digest=sha256:b938b244b4013c47a0890cafb4c5d9f4374c1e8fdf823d90686de385e532ef06

Observation b8e90943-66f2-42a2-9cd8-2fee265ddf93 · outbound

This paper cites Conservative set valued fields, automatic differentiation, stochastic gradient methods and deep learning.

Mathematical analysis of the gradients in deep learning Conservative set valued fields, automatic differentiation, stochastic gradient methods and deep learning

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T14:10:56.212131Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T14:10:55.101003Z digest=sha256:196fbaa8ee140822c1ed4ad6f0f42d330a97ebfd06982c8dc68d93b8b9352989

Observation eec2a3c5-b9d9-47db-bffd-d581aebd50e8 · outbound

This paper cites Differentiating nonsmooth solutions to parametric monotone inclusion problems.

Mathematical analysis of the gradients in deep learning Differentiating nonsmooth solutions to parametric monotone inclusion problems

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T14:10:56.199136Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T14:10:55.104409Z digest=sha256:a7bd350d7f2b79738cb685679e5dd6cbcf5ce6cec8c4478e53140fdbba193207

Observation 7de81cdc-eaaf-45fa-adcf-190e8cac1a6e · outbound

This paper cites Automatic differentiation of nonsmooth iterative algorithms.

Mathematical analysis of the gradients in deep learning Automatic differentiation of nonsmooth iterative algorithms

Reference 9

Resolution
verified exact
local_arxiv, observed 2026-08-10T14:10:55.865319Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T14:10:55.108574Z digest=sha256:4199d204f293740f3cccbb0850df4043fd9f9c4640df5e1f11120be0b7fe9540

Observation 981b5c36-5e7a-45cf-9ba0-d2ec49606f72 · outbound

This paper cites A proof of conver- gence for gradient descent in the training of artificial neur al networks for constant target functions.

Mathematical analysis of the gradients in deep learning A proof of conver- gence for gradient descent in the training of artificial neur al networks for constant target functions

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T14:10:56.186067Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T14:10:55.112707Z digest=sha256:f78d976dc0296d905767992abaa64bbf05daa5cc3824e9e1f15284b3a9806cd1

Observation 67f90682-0a80-4db7-96cc-8f883e6cd23d · outbound

This paper cites Non-convergence of stochastic gradient descent in the training of deep neural networks.

Mathematical analysis of the gradients in deep learning Non-convergence of stochastic gradient descent in the training of deep neural networks

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T14:10:56.173592Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T14:10:55.116780Z digest=sha256:45f5739e6e3c827ae65a9a686165b2487a929f8a5cf8ae18e265969711f81648

Observation 96174e1b-66a5-43e1-9f10-caec25de07ed · outbound

This paper cites Landscape analysis for shallow neural networks: complete classification of critical point s for affine target functions.

Mathematical analysis of the gradients in deep learning Landscape analysis for shallow neural networks: complete classification of critical point s for affine target functions

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T14:10:56.161634Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T14:10:55.121575Z digest=sha256:6833078ed30ecf18963ad082a9bfab6a03de3a9e17aee4b74cc793fb8dcccdad

Observation 298ebf6c-13f7-4d44-af12-ecf75480c4d6 · outbound

This paper cites On the mathematical foundations of learning.

Mathematical analysis of the gradients in deep learning On the mathematical foundations of learning

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T14:10:56.149980Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T14:10:55.125374Z digest=sha256:7601c7864d94914c2077b521fd11773e759faa79b4e6b75fc57514c32f676c19

Observation c4d9f116-ace4-4096-9460-39d1965ca551 · outbound

This paper cites Conservative and semismooth derivatives are equiv- alent for semialgebraic maps.

Mathematical analysis of the gradients in deep learning Conservative and semismooth derivatives are equiv- alent for semialgebraic maps

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T14:10:56.137872Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T14:10:55.128887Z digest=sha256:1c6b76ccaaf5e0026b8d0494d354fc9a7d5a0cf8518882eb569d412644df9084

Observation 4dbe2bae-0527-4de1-a815-73e8aa838dae · outbound

This paper cites an unresolved cited work.

Mathematical analysis of the gradients in deep learning Unresolved cited work

Reference 15

Resolution
unresolved
raw_fallback, observed 2026-08-10T14:10:56.125431Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T14:10:55.132991Z digest=sha256:f088dcb847318d46de48b2a3c63c5b43ed8db8023b888c3b56d2c52edc28dab3

Observation f9aaa7f2-7500-4db6-b94c-683116f86386 · outbound

This paper cites Non-convergence of Adam and other adaptive stochastic gradient descent optimization methods for non-vanishing learning rates.

Mathematical analysis of the gradients in deep learning Non-convergence of Adam and other adaptive stochastic gradient descent optimization methods for non-vanishing learning rates

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-10T14:10:55.137201Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T14:10:55.137201Z digest=sha256:e9bf5e6d98d9420c6000737c99fc83269448cce985959c0fb4b7412328d12b76

Observation c4232203-5f03-46fa-aed2-d4310676e10d · outbound

This paper cites Convergence of stochastic gradient descent schemes for Lojasiewicz-landscapes.

Mathematical analysis of the gradients in deep learning Convergence of stochastic gradient descent schemes for Lojasiewicz-landscapes

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-10T14:10:55.140914Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T14:10:55.140914Z digest=sha256:c1f2a9c144985f86253f55432bce4cc3d7285c8977a83d175fd3712a53ffb814

Observation 8a76bc57-5b4e-49be-acea-7ad9090d1fc4 · outbound

This paper cites Adam-family Methods with Decoupled Weight Decay in Deep Learning.

Mathematical analysis of the gradients in deep learning Adam-family Methods with Decoupled Weight Decay in Deep Learning

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-10T14:10:55.144839Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T14:10:55.144839Z digest=sha256:1361e71fd4af0d1ef147db9fce678562b1d35f22e654a0f20257d412c4632308

Observation 45dc87df-2ace-4e74-80f8-32e5a802df13 · outbound

This paper cites Sub-Optimal Local Minima Exist for Neural Networks with Almost All Non-Linear Activations.

Mathematical analysis of the gradients in deep learning Sub-Optimal Local Minima Exist for Neural Networks with Almost All Non-Linear Activations

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-10T14:10:55.149307Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T14:10:55.149307Z digest=sha256:bb4719f0a33f33e6137a1ef9260f2a8892ece5a318c0d265d7765eefe42e8d38

Observation 237b81ec-86ee-4ab1-a3c2-a32df0fd1f55 · outbound

This paper cites Towards a Mathematical Understanding of Neural Network-Based Machine Learning: what we know and what we don't.

Mathematical analysis of the gradients in deep learning Towards a Mathematical Understanding of Neural Network-Based Machine Learning: what we know and what we don't

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-10T14:10:55.153180Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T14:10:55.153180Z digest=sha256:cdcd1cff4613583c2a4f389324089ae8eb4bb0b10bdddc04abd2b96dc612af9f

Observation 0b27e058-fa9f-4ea2-a1b0-eacf244f4589 · outbound

This paper cites an unresolved cited work.

Mathematical analysis of the gradients in deep learning Unresolved cited work

Reference 21

Resolution
unresolved
raw_fallback, observed 2026-08-10T14:10:56.112630Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T14:10:55.157102Z digest=sha256:f8347d252d5df1a4828555b040b8a22f84b22c37ee089ad0323e6cabfae7fe2e

Observation d3d9e9cc-cec9-4211-a353-d84650cddd89 · outbound

This paper cites Blow up phenomena for gradient descent optimization methods in the training of artificial neural networks.

Mathematical analysis of the gradients in deep learning Blow up phenomena for gradient descent optimization methods in the training of artificial neural networks

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-10T14:10:55.160595Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T14:10:55.160595Z digest=sha256:925ab7b141fcca6076e2a5ccf24b34fbd80fc20c09955ff3f2fccbc36db300c4

Observation 01304cc6-e5b7-404b-a619-5801727b4eb9 · outbound

This paper cites Handbook of Convergence Theorems for (Stochastic) Gradient Methods.

Mathematical analysis of the gradients in deep learning Handbook of Convergence Theorems for (Stochastic) Gradient Methods

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-10T14:10:55.164695Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T14:10:55.164695Z digest=sha256:13ec606260db0772fbabd721af7846f386e806eacecd6748b55b497d81afdc38

Observation 34356b5b-423a-401e-be69-5d66b7e03ae3 · outbound

This paper cites Approximation results for Gradient Descent trained Shallow Neural Networks in $1d$.

Mathematical analysis of the gradients in deep learning Approximation results for Gradient Descent trained Shallow Neural Networks in $1d$

Reference 24

Resolution
verified exact
local_arxiv, observed 2026-08-10T14:10:55.769225Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T14:10:55.168292Z digest=sha256:a9855ae509fffe145e8abfe321545141e705e023fe4864e5c64b4cc5ae61a8ec

Observation 7d58f89a-2b2c-4fde-92c0-aca6e4701d44 · outbound

This paper cites Non-convergence to global minimizers in data driven supervised deep learning: Adam and stochastic gradient descent optimization provably fail to converge to global minimizers in the training of deep neural networks with ReLU activation.

Mathematical analysis of the gradients in deep learning Non-convergence to global minimizers in data driven supervised deep learning: Adam and stochastic gradient descent optimization provably fail to converge to global minimizers in the training of deep neural networks with ReLU activation

Reference 25

Resolution
verified exact
local_arxiv, observed 2026-08-10T14:10:55.752584Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T14:10:55.171951Z digest=sha256:9198a1b97ab95596e742d59c0a5b8b974d7c6c6eff91d3ee3866343afe25ebaf

Observation 65143bab-0e87-4d47-b350-06063f650409 · outbound

This paper cites Convergence proof for stochastic gradient descent in the training of deep neural networks with ReLU activation for constant target functions.

Mathematical analysis of the gradients in deep learning Convergence proof for stochastic gradient descent in the training of deep neural networks with ReLU activation for constant target functions

Reference 26

Resolution
verified exact
local_arxiv, observed 2026-08-10T14:10:55.737848Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T14:10:55.176028Z digest=sha256:6cc12ca1a1a19c3c32a1ecb80f39d9b678818ee6fd48a8728d9ece5d17428aa1

Observation af440117-6d22-4089-a733-751858c2023f · outbound

This paper cites Convergence to good non-optimal critical points in the training of neural networks: Gradient descent optimization with one random initialization overcomes all bad non-global local minima with high probability.

Mathematical analysis of the gradients in deep learning Convergence to good non-optimal critical points in the training of neural networks: Gradient descent optimization with one random initialization overcomes all bad non-global local minima with high probability

Reference 27

Resolution
verified exact
local_arxiv, observed 2026-08-10T14:10:55.721846Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T14:10:55.180503Z digest=sha256:a7e5289170c908c4b4e51c96ed4f8b27c3fa985b1d9f3aaf4f12faf08189a49a

Observation 23c3f78c-f589-4871-bc95-8213a9807e5f · outbound

This paper cites Mathematical Introduction to Deep Learning: Methods, Implementations, and Theory.

Mathematical analysis of the gradients in deep learning Mathematical Introduction to Deep Learning: Methods, Implementations, and Theory

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-10T14:10:55.184143Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T14:10:55.184143Z digest=sha256:1110c733b360208c289c6d153fc606ab646516ec15aa4a96abde599a17c15956

Observation 56865051-e276-4cac-8445-25a6fb17bf8d · outbound

This paper cites On the existence of global minima and convergence analyses for gradient descent methods in the training of deep neural networks.

Mathematical analysis of the gradients in deep learning On the existence of global minima and convergence analyses for gradient descent methods in the training of deep neural networks

Reference 29

Resolution
verified exact
local_arxiv, observed 2026-08-10T14:10:55.693711Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T14:10:55.187721Z digest=sha256:aa17e93ad85902e206872fdd4a1c38a83dc9c01891cc8addff99e1c7fcb6e4fe

Observation 77ca3d51-68d6-42d2-9cb2-b1b4cb6e0829 · outbound

This paper cites On the existence of global minima and convergence analyses for gradient descent methods in the training of dee p neural networks.

Mathematical analysis of the gradients in deep learning On the existence of global minima and convergence analyses for gradient descent methods in the training of dee p neural networks

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T14:10:56.101057Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T14:10:55.191988Z digest=sha256:3989c0a1d3337b8193507f4808c38e8591c1241885e24b315b47bdab32cfd02b

Observation 730c4d9f-1152-4a86-943e-3dc45a151caa · outbound

This paper cites A proof of convergence for stochastic gradient descent in the training of artificial neural networks with ReLU activation for constant target functions.

Mathematical analysis of the gradients in deep learning A proof of convergence for stochastic gradient descent in the training of artificial neural networks with ReLU activation for constant target functions

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T14:10:56.089558Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T14:10:55.196026Z digest=sha256:b8af959ae0c5b815c8f0fdc4d5c70e19e4fa7b3858466bbe9b78553d47328e18

Observation 4766fd9e-0563-4699-8fc0-cedf26672fda · outbound

This paper cites Convergence analysis for gradient flows in the training of artificial neural networks with ReLU activation.

Mathematical analysis of the gradients in deep learning Convergence analysis for gradient flows in the training of artificial neural networks with ReLU activation

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T14:10:56.076929Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T14:10:55.200155Z digest=sha256:8b031841b76a62d0ba9aecfda99c58f7513e81e1af2dc42a5069d813dfceda3f

Observation 4eedb1c9-c93b-4322-95cf-0880eaeeb427 · outbound

This paper cites Non-convergence to global minimizers for Adam and stochastic gradient descent optimization and constructions of local minimizers in the training of artificial neural networks.

Mathematical analysis of the gradients in deep learning Non-convergence to global minimizers for Adam and stochastic gradient descent optimization and constructions of local minimizers in the training of artificial neural networks

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-10T14:10:55.203554Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T14:10:55.203554Z digest=sha256:2983199e5165bfd26bdf83d61f6e189fc37e290c17f408ade45fe15c619f1623

Observation 69f09e38-d5c9-4b34-a18e-2ce0ddbe999e · outbound

This paper cites M., and Lee, J.

Mathematical analysis of the gradients in deep learning M., and Lee, J

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T14:10:56.064425Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T14:10:55.207272Z digest=sha256:7ef5dbdd3890a375b5dd9acd62017269fba41c6487a90c2f44d53439bb2156ef

Observation 34629051-9f5a-49aa-8680-06075f1cd961 · outbound

This paper cites Does a sparse ReLU network training problem always admit an optimum?.

Mathematical analysis of the gradients in deep learning Does a sparse ReLU network training problem always admit an optimum?

Reference 35

Resolution
metadata mismatch
local_arxiv, observed 2026-08-10T14:10:55.664692Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T14:10:55.211624Z digest=sha256:5d4e141d952e1afdd03aded334454f76de863f3a57029f8e85830d4cf0611497

Observation fee10728-e53e-4f3a-9771-574812d91ad8 · outbound

This paper cites On correctness of automatic differentiation for non-differentiable functions.

Mathematical analysis of the gradients in deep learning On correctness of automatic differentiation for non-differentiable functions

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T14:10:56.052008Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T14:10:55.215356Z digest=sha256:428a8d5e95afaa61bea7defa392765c3956754e0afe0fa8681ef901820bd152e

Observation 59c9386c-576d-46db-b028-18257b59b67d · outbound

This paper cites S., and Tian, T.

Mathematical analysis of the gradients in deep learning S., and Tian, T

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T14:10:56.040243Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T14:10:55.219333Z digest=sha256:0ab0f6c7f6dd29fd995543a118c0a916c990c266f231161ef53726ecd9ae6b6f

Observation a2a99617-256c-44a6-ac47-bcdbedb122a8 · outbound

This paper cites an unresolved cited work.

Mathematical analysis of the gradients in deep learning Unresolved cited work

Reference 38

Resolution
unresolved
raw_fallback, observed 2026-08-10T14:10:56.028119Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T14:10:55.223509Z digest=sha256:e39b1155718cfee6c81e5cdf16158a14ee1877a2846ee04a1b5a4f63607676aa

Observation d9636a0d-b071-4e95-835d-0f5b0a7940b7 · outbound

This paper cites Introductory lectures on convex optimization , vol.

Mathematical analysis of the gradients in deep learning Introductory lectures on convex optimization , vol

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T14:10:56.015946Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T14:10:55.227077Z digest=sha256:004c3bad218b009cb0be63e02096b253134179c61799bf98cab2e8eb8ef6a5b8

Observation e73bedc5-766f-4a3e-9074-5399620c1a26 · outbound

This paper cites Automatic differentiation in PyTorch.

Mathematical analysis of the gradients in deep learning Automatic differentiation in PyTorch

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T14:10:56.003282Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T14:10:55.230372Z digest=sha256:3094425c659c6b6c7c2219f207b665c0ef5cde7df41168af669540aca6107077

Observation 18600bc8-8a6f-42d7-acfd-8744b14e2c17 · outbound

This paper cites PyTorch: An Imperative Style, High-Performance Deep Learning Library.

Mathematical analysis of the gradients in deep learning PyTorch: An Imperative Style, High-Performance Deep Learning Library

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-10T14:10:55.234656Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T14:10:55.234656Z digest=sha256:3576b63c6071bc99a090cd41eabb6979da17706db900aa880dd147d8ebb2a585

Observation a3d61f3a-0a6d-45ac-8add-0c11fabcb32f · outbound

This paper cites Conservative parametric optimality and the ridge method for tame min-max problems.

Mathematical analysis of the gradients in deep learning Conservative parametric optimality and the ridge method for tame min-max problems

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T14:10:55.991827Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T14:10:55.238390Z digest=sha256:f07e8ae82a8734bd417a491d606c83175a48f63195d021459b218efefca5aded

Observation c5e849ba-d9cf-4a7e-81b3-1e073f726c5d · outbound

This paper cites Topological properties of the set of functions generated by neural networks of fixed size.

Mathematical analysis of the gradients in deep learning Topological properties of the set of functions generated by neural networks of fixed size

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T14:10:55.980143Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T14:10:55.242309Z digest=sha256:4563fb7001057ff57465c2aa696788297b5ac4e267376944530465e411cfbf0f

Observation d8bd2e87-a613-4e51-8aa5-9ae1eee4fbcc · outbound

This paper cites Mathematical theory of deep learning.

Mathematical analysis of the gradients in deep learning Mathematical theory of deep learning

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-10T14:10:55.246396Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T14:10:55.246396Z digest=sha256:d678e9fcc0515bc21b1846fa33d962ddd0e361e2ecc429e517867fedd553d525

Observation 04fab9aa-cc6d-4de7-928d-caf7c839d88a · outbound

This paper cites On the Convergence of Adam and Beyond.

Mathematical analysis of the gradients in deep learning On the Convergence of Adam and Beyond

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-10T14:10:55.250440Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T14:10:55.250440Z digest=sha256:efa11570973d434be59d27b49a441be9d6d7b4e0458515b8d0dba0dd0d6e416c

Observation 6596ba1a-4111-4321-a318-7bb990f3c059 · outbound

This paper cites T., and Wets, R.

Mathematical analysis of the gradients in deep learning T., and Wets, R

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T14:10:55.967280Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T14:10:55.254360Z digest=sha256:1dbeee2d11ee8238bb9e3a1d49692848bc0804d6944f5b2bb8114cfc11e1a964

Observation 587b5c32-04c1-459a-b2e3-85e6a28f2179 · outbound

This paper cites An overview of gradient descent optimization algorithms.

Mathematical analysis of the gradients in deep learning An overview of gradient descent optimization algorithms

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-10T14:10:55.258087Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T14:10:55.258087Z digest=sha256:6cf79a3219fc5c26ade23c813977d7e535d73ebd60bc8e78c81b9be640c65668

Observation 3e73331b-66c2-4d4e-a9b9-b4224f0c7db5 · outbound

This paper cites Spurious Local Minima are Common in Two-Layer ReLU Neural Networks.

Mathematical analysis of the gradients in deep learning Spurious Local Minima are Common in Two-Layer ReLU Neural Networks

Reference 48

Resolution
verified exact
local_arxiv, observed 2026-08-10T14:10:55.559559Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T14:10:55.261614Z digest=sha256:46ea5eabccbd2156d228f1ffc0b041a51e9627d62fc1d4074b867f7cec5d6790

Observation aa99968d-afbc-48fa-8857-030d520112f9 · outbound

This paper cites The gradient’s limit of a definable family of functions is a co nservative set-valued field.

Mathematical analysis of the gradients in deep learning The gradient’s limit of a definable family of functions is a co nservative set-valued field

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-10T14:10:55.265908Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T14:10:55.265908Z digest=sha256:db51a9d1e8781297de7d177bd2a0be2d0ad59e26114d049e6bb31cbe17189c05

Observation 72b69994-0986-4f7c-8459-8d7434480bc8 · outbound

This paper cites an unresolved cited work.

Mathematical analysis of the gradients in deep learning Unresolved cited work

Reference 50

Resolution
unresolved
raw_fallback, observed 2026-08-10T14:10:55.956109Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T14:10:55.269590Z digest=sha256:fe1cc92275bfd0ae45f2f40fdd8955de50e70e0e989e766f39458ee23e311829

Observation 26fff830-1fc2-4193-a377-d46ca9e8f0c2 · outbound

This paper cites Optimization for deep learning: theory and algorithms.

Mathematical analysis of the gradients in deep learning Optimization for deep learning: theory and algorithms

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-10T14:10:55.368199Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T14:10:55.368199Z digest=sha256:490603f8e6b69844c6ef5a294e79833df75caaaa90bd0d653b2ab149c0f59fca

Observation f61b343a-98b8-488c-98a2-f0f5034bd062 · outbound

This paper cites Local minima in training of neural networks.

Mathematical analysis of the gradients in deep learning Local minima in training of neural networks

Reference 52

Resolution
verified exact
local_arxiv, observed 2026-08-10T14:10:55.470243Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T14:10:55.372632Z digest=sha256:ac56b8cea4419089b260ec12b0875e36f68fefe2aa55410b32321298870acdbe

Observation 2e04d224-a0a5-4b26-8625-9e01e5213426 · outbound

This paper cites S., and Bruna, J.

Mathematical analysis of the gradients in deep learning S., and Bruna, J

Reference 53

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T14:10:55.944838Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T14:10:55.376505Z digest=sha256:d52040c70776d3fd2527f3d8308b643db162c7007ac991a1abc57187e7488874

Observation c50e8cdc-96f1-4cd6-b8f4-348c3b7ac367 · outbound

This paper cites Approximation and Gradient Descent Training with Neural Networks.

Mathematical analysis of the gradients in deep learning Approximation and Gradient Descent Training with Neural Networks

Reference 54

Resolution
verified exact
local_arxiv, observed 2026-08-10T14:10:55.453704Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T14:10:55.379861Z digest=sha256:f14a9e72e78a13b656d33b204750cc5f217c91c20765a66123c8710ed1153511

Observation b09235ce-e803-4d66-828b-d312a10535e8 · outbound

This paper cites Approximation results for gradient flow trained neural netw orks.

Mathematical analysis of the gradients in deep learning Approximation results for gradient flow trained neural netw orks

Reference 55

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T14:10:55.933568Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T14:10:55.383559Z digest=sha256:a6214b53438443d6058fa1c4d98f3b56e074e093b93acea807fd0b864af312e6

Observation 18d67b15-0b65-4b23-84f4-71b7548b57a3 · outbound

This paper cites Adam-family Methods for Nonsmooth Optimization with Convergence Guarantees.

Mathematical analysis of the gradients in deep learning Adam-family Methods for Nonsmooth Optimization with Convergence Guarantees

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-10T14:10:55.387153Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T14:10:55.387153Z digest=sha256:bb6f324b8a173b6bbcef4724039944feaf0a963d2ad58ec462fff13191ae424a

Observation 409d5b73-5735-41b6-b555-24d8cf6ce3d2 · outbound

This paper cites Stochastic Subgradient Methods with Guaranteed Global Stability in Nonsmooth Nonconvex Optimization.

Mathematical analysis of the gradients in deep learning Stochastic Subgradient Methods with Guaranteed Global Stability in Nonsmooth Nonconvex Optimization

Reference 57

Resolution
unresolved
no resolver link, observed 2026-08-10T14:10:55.390854Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T14:10:55.390854Z digest=sha256:a1a7fd1736071cb2ec64df8c341a3af74a877cbbed581289efd03d26d14134e9

Observation 02b81597-8b7b-4ff3-98c1-b4846b707d05 · outbound

This paper cites an unresolved cited work.

Mathematical analysis of the gradients in deep learning Unresolved cited work

Reference 58

Resolution
unresolved
raw_fallback, observed 2026-08-10T14:10:55.922227Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T14:10:55.394691Z digest=sha256:9c67d586855d8c729d50a668cc35bca872546ada15361046671d56396025b844

Pith citing papers

No inbound Pith citation observations are available.