Pith. sign in

Paper Citation Record · LEDGER

Large scale distributed neural network training through online distillation

As of 23 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 21 inbound Pith citation observations for arXiv:1804.03235.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
1804.03235 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 21 of 21 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-23T06:30:58.430688+00:00

measured 21 of 21 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-15T21:53:22.089226Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

152
arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 97e4cdf4-ac17-452c-9363-cfb7f81f6840 · inbound

Distilling On-Device Intelligence at the Network Edge cites this paper.

Distilling On-Device Intelligence at the Network Edge Large scale distributed neural network training through online distillation

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-14T13:06:33.897510Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T13:06:33.897510Z digest=sha256:efddf5b816cf8aea21df6ec16bef32fbbb8b8f8fbbdd532b4c5095deade0850f

Observation 00f6897a-b2ae-4325-b093-c99bc0e223a5 · inbound

Learning Representations and Agents for Information Retrieval cites this paper.

Learning Representations and Agents for Information Retrieval Large scale distributed neural network training through online distillation

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-14T13:00:57.500008Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T13:00:57.500008Z digest=sha256:277d3b8970f0fcdc85d4250a19d48c947a01d33b27f9783a790690ca10b45238

Observation b4290041-96f0-4152-9a82-96f990142131 · inbound

Emerging Properties in Self-Supervised Vision Transformers cites this paper.

Emerging Properties in Self-Supervised Vision Transformers Large scale distributed neural network training through online distillation

Reference 1

Resolution
verified exact
arxiv_id, observed 2026-05-16T14:04:51.498387Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-05-16T14:04:51.458382Z digest=sha256:1b2c49621c7e25b68a8e9022775c86ff2e3136fc41d06463e659ff7b116093b8

Observation b717688e-bec4-43c8-9fe5-b5d3e8375e3b · inbound

Vision Transformers Need Registers cites this paper.

Vision Transformers Need Registers Large scale distributed neural network training through online distillation

Reference 182

Resolution
metadata mismatch
arxiv_id, observed 2026-05-13T09:41:38.167514Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-05-13T09:41:37.937046Z digest=sha256:58ac020f433ea1f87bf0a5664217b65451c9b70beeab7d7ef36418e58ce63886

Observation ce14800f-1557-42fb-87fc-dd9ac110c15d · inbound

Gemma 2: Improving Open Language Models at a Practical Size cites this paper.

Gemma 2: Improving Open Language Models at a Practical Size Large scale distributed neural network training through online distillation

Reference 159

Resolution
metadata mismatch
arxiv_id, observed 2026-05-10T12:11:16.544210Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-05-10T12:11:16.326752Z digest=sha256:a116cea5b7869bced77c17ab3e25462ba8d76c03df645390349df481510b877f

Observation 51dd6677-6aa8-4775-abfc-f5a76203992e · inbound

A Collaborative Ensemble Framework for CTR Prediction cites this paper.

A Collaborative Ensemble Framework for CTR Prediction Large scale distributed neural network training through online distillation

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-12T16:17:31.268231Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T16:17:31.268231Z digest=sha256:1153d939aec163b1bd92081cb05961f8712e0729f99726b689e5cf41e5c33510

Observation 1b6250d1-edef-448b-a617-b70536537829 · inbound

Beyond Model Scale Limits: End-Edge-Cloud Federated Learning with Self-Rectified Knowledge Agglomeration cites this paper.

Beyond Model Scale Limits: End-Edge-Cloud Federated Learning with Self-Rectified Knowledge Agglomeration Large scale distributed neural network training through online distillation

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-10T22:48:56.719676Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T22:48:56.719676Z digest=sha256:a2a0a273fd48de2c346fe7fae58efcdfda3423a10757709137ae8e71ba4f25ab

Observation 23fca9e2-d940-40cc-919f-24fd1e71bf8e · inbound

Gemma 3 Technical Report cites this paper.

Gemma 3 Technical Report Large scale distributed neural network training through online distillation

Reference 2

Resolution
verified exact
arxiv_id, observed 2026-05-22T22:22:12.260167Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-05-22T22:18:55.976503Z digest=sha256:312b3c3c7b95cad3a22510b4cc9186e42f678fc952328b99f898122a11aad320

Observation c2707762-2e1a-4b25-8c84-cb9a488a9edc · inbound

Modular Federated Learning: A Meta-Framework Perspective cites this paper.

Modular Federated Learning: A Meta-Framework Perspective Large scale distributed neural network training through online distillation

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-15T21:53:22.089226Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:53:22.089226Z digest=sha256:03b787581fb28f78254983dc61c30c87e4ebc1ab8a5296860521fe50e3a92b90

Observation d8b87f14-6a0b-4696-b8be-4cb61f993b47 · inbound

Gemini 2.5: Pushing the Frontier with Advanced Reasoning, Multimodality, Long Context, and Next Generation Agentic Capabilities cites this paper.

Gemini 2.5: Pushing the Frontier with Advanced Reasoning, Multimodality, Long Context, and Next Generation Agentic Capabilities Large scale distributed neural network training through online distillation

Reference 1

Resolution
verified exact
arxiv_id, observed 2026-05-19T05:52:07.665258Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-05-19T05:48:02.828938Z digest=sha256:de064441bc3390b79b8aebceadef661a9a8b7712eac7f0f9a8b1045dba5ab4b9

Observation c4c0afed-2076-466a-a5b2-92dd9baf6b6c · inbound

Optimal Transceiver Design in Over-the-Air Federated Distillation cites this paper.

Optimal Transceiver Design in Over-the-Air Federated Distillation Large scale distributed neural network training through online distillation

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-06T15:43:46.990439Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:43:46.990439Z digest=sha256:5b0e2882d82d0df03966ac01fb2771c538e2ec88b590a7547717cf125ace8c75

Observation 2b55ff7a-7fd7-4257-9a7e-274e60bbda3f · inbound

The Ratchet Effect in Silico: How Interaction Drives Cumulative Intelligence in Large Language Models cites this paper.

The Ratchet Effect in Silico: How Interaction Drives Cumulative Intelligence in Large Language Models Large scale distributed neural network training through online distillation

Reference 6

Resolution
verified exact
arxiv_id, observed 2026-05-19T02:42:00.002358Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-05-19T02:38:27.510872Z digest=sha256:7bd09ecc557c40763e8b827d9a57540b8f1889f1c49bca504e4829f9e20cd173

Observation 95e477ed-9046-42c0-96bd-44e82e0536af · inbound

HypEHR: Hyperbolic Modeling of Electronic Health Records for Efficient Question Answering cites this paper.

HypEHR: Hyperbolic Modeling of Electronic Health Records for Efficient Question Answering Large scale distributed neural network training through online distillation

Reference 178

Resolution
metadata mismatch
arxiv_id, observed 2026-05-09T23:54:45.166952Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-05-09T23:51:47.724033Z digest=sha256:69c6bdd01f83c551fee881699a813f663944e73e1b72c66bcad58d50ac030b1b

Observation 2c0d22fa-eedf-4fff-adbc-f7a777f31145 · inbound

LatentBurst: A Fast and Efficient Multi Frame Super-Resolution for Hexadeca-Bayer Pattern CIS images cites this paper.

LatentBurst: A Fast and Efficient Multi Frame Super-Resolution for Hexadeca-Bayer Pattern CIS images Large scale distributed neural network training through online distillation

Reference 25

Resolution
verified exact
arxiv_id, observed 2026-05-11T20:31:11.968340Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-05-08T08:48:09.398221Z digest=sha256:07d03edf9d92490d917a810e7ddf463dbbea942d0614582f5365cbc54f017537

Observation 59aebcfc-3205-4747-9a21-72a29377f62c · inbound

Enabling Federated Inference via Unsupervised Consensus Embedding cites this paper.

Enabling Federated Inference via Unsupervised Consensus Embedding Large scale distributed neural network training through online distillation

Reference 17

Resolution
verified exact
arxiv_id, observed 2026-05-11T18:36:07.875551Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-05-08T15:06:27.437479Z digest=sha256:a7694ddda722319d2ecea41ffcf73b0fab02020c0f04dfe95c17e18c05b60e9c

Observation 65ed5b5e-5cc9-4a64-9119-4dc667e6c3ef · inbound

Function-Space ADMM for Decentralized Federated Learning: A Control Theoretic Perspective cites this paper.

Function-Space ADMM for Decentralized Federated Learning: A Control Theoretic Perspective Large scale distributed neural network training through online distillation

Reference 22

Resolution
verified exact
arxiv_id, observed 2026-05-12T06:36:26.935644Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-05-12T04:07:54.716306Z digest=sha256:a1718611e40c52bcea2bbe959b06d9d8a7ae6be7fe05e6876bc5032cbe543aa3

Observation 7ba269ad-7474-4973-a851-add08194d938 · inbound

Optimized Federated Knowledge Distillation with Distributed Neural Architecture Search cites this paper.

Optimized Federated Knowledge Distillation with Distributed Neural Architecture Search Large scale distributed neural network training through online distillation

Reference 35

Resolution
verified exact
arxiv_id, observed 2026-05-21T05:43:58.955772Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-05-21T05:42:32.050913Z digest=sha256:4586e82cf7d55a728a10b1a3afd6aded14eabaf05734e5d47c3d37a41b88149d

Observation f719f77c-941f-4a0f-8cc4-af0a02c5fb75 · inbound

Scaling Laws for Task-Specific LLM Distillation cites this paper.

Scaling Laws for Task-Specific LLM Distillation Large scale distributed neural network training through online distillation

Reference 18

Resolution
verified exact
arxiv_id, observed 2026-07-04T17:40:00.752562Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-06-25T23:29:49.787477Z digest=sha256:bc9e789a0f4f94de5f6d4605c2d7234f6e79a02eb282c818cf5a163747cf791a

Observation 82af35fa-8f92-4ab7-93cb-f7edf6ec8898 · inbound

Bridging Compute- and Data-Optimal Pretraining cites this paper.

Bridging Compute- and Data-Optimal Pretraining Large scale distributed neural network training through online distillation

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-01T03:02:01.645052Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T03:02:01.645052Z digest=sha256:456c83b8882689b26c3d33f4d65f3391f26fedd6af797be6474c85642fc99b87

Observation 82fd8473-1d5f-43ee-9503-fddb8d10444f · inbound

Slot2Text: Object-Centric Visual Tokenization for Efficient and Spatially Traceable Surgical MLLMs cites this paper.

Slot2Text: Object-Centric Visual Tokenization for Efficient and Spatially Traceable Surgical MLLMs Large scale distributed neural network training through online distillation

Reference 116

Resolution
unresolved
no resolver link, observed 2026-08-06T00:10:37.428710Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T00:10:37.428710Z digest=sha256:5d6108a2be76f0e13950351b76981fbeb301253298e365a34129b8c2a47ba011

Observation 49166915-6c14-4ccc-8d84-dad375d7ec9c · inbound

Out-of-Distribution Federated Distillation with Domain-Aware Proxy cites this paper.

Out-of-Distribution Federated Distillation with Domain-Aware Proxy Large scale distributed neural network training through online distillation

Reference 2022

Resolution
unresolved
no resolver link, observed 2026-08-14T04:37:52.575322Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T04:37:52.575322Z digest=sha256:e51b8c10429e7675ab49c7f930166f4d467c80299c301bac3c3f0fa8dc62c14c