Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T04:20:04.267992Z
Paper Citation Record · LEDGER
As of 7 August 2026, this Paper Citation Record lists 31 of 31 outbound references and 5 inbound Pith citation observations for arXiv:2506.10911.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T04:20:04.267992Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-07-12T12:16:10.904456Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-07-03T04:07:37.154141Z
31 of 31 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 17186e22-de31-4b7b-80cc-87472de0c03a · outbound
NoLoCo: No-all-reduce Low Communication Training Method for Large Models Unresolved cited work
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation b44aafe4-29b1-4f1c-b8fb-76420366909f · outbound
NoLoCo: No-all-reduce Low Communication Training Method for Large Models Communication-Efficient Language Model Training Scales Reliably and Robustly: Scaling Laws for DiLoCo
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation dc4ae879-c1d3-40eb-9951-968107914fb6 · outbound
NoLoCo: No-all-reduce Low Communication Training Method for Large Models Efficient Training of Large Language Models on Distributed Infrastructures: A Survey
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6d204c69-99c5-4a21-bd5b-525ec5c9bf7f · outbound
NoLoCo: No-all-reduce Low Communication Training Method for Large Models Accelerating Large Language Model Training with 4D Parallelism and Memory Consumption Estimator
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3e815e8f-b7d8-4f4f-a5e3-437ea48638f0 · outbound
NoLoCo: No-all-reduce Low Communication Training Method for Large Models Multi-modal retrieval for large language model based speech recognition
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation b01983a9-bd29-40e3-9935-155b55572131 · outbound
NoLoCo: No-all-reduce Low Communication Training Method for Large Models The Llama 3 Herd of Models
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 77259df9-ea19-4f21-ac1c-ae6023676458 · outbound
NoLoCo: No-all-reduce Low Communication Training Method for Large Models Gossip learning as a decentralized alternative to federated learning
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 3374d221-5816-47ef-83dd-ecb16684645c · outbound
NoLoCo: No-all-reduce Low Communication Training Method for Large Models INTELLECT-1 Technical Report
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3c9ed00b-a904-4091-89e0-3e9c0ab20b07 · outbound
NoLoCo: No-all-reduce Low Communication Training Method for Large Models Eager Updates For Overlapped Communication and Computation in DiLoCo
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b4b1e656-90ed-4aec-8020-cb5f269abe2c · outbound
NoLoCo: No-all-reduce Low Communication Training Method for Large Models Branch-Train-Merge: Embarrassingly Parallel Training of Expert Language Models
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 10e37934-bee7-490a-a7d6-0280c0aea2cc · outbound
NoLoCo: No-all-reduce Low Communication Training Method for Large Models DeepSeek-V2: A Strong, Economical, and Efficient Mixture-of-Experts Language Model
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 530c52cb-cc8e-42ee-8693-b14682ecb157 · outbound
NoLoCo: No-all-reduce Low Communication Training Method for Large Models Ring Attention with Blockwise Transformers for Near-Infinite Context
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a4cc0a39-b71e-4342-829e-77b1049da533 · outbound
NoLoCo: No-all-reduce Low Communication Training Method for Large Models Voxtlm: Unified decoder-only models for consolidating speech recognition, synthesis and speech, text continuation tasks
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation aa96d7c7-d788-4ba7-87a7-e7b9fa5ab439 · outbound
NoLoCo: No-all-reduce Low Communication Training Method for Large Models Decoupled momentum optimization.arXiv preprint arXiv:2411.19870,
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 95c78d5c-61c1-4da5-ab96-fd94ef911029 · outbound
NoLoCo: No-all-reduce Low Communication Training Method for Large Models AudioPaLM: A Large Language Model That Can Speak and Listen
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 32e693ef-cd80-4b41-b600-8517c3f41108 · outbound
NoLoCo: No-all-reduce Low Communication Training Method for Large Models Seq1F1B: Efficient Sequence-Level Pipeline Parallelism for Large Language Model Training
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cd3950f5-67fb-440c-b0d9-34d6f7f6243a · outbound
NoLoCo: No-all-reduce Low Communication Training Method for Large Models Gemini: A Family of Highly Capable Multimodal Models
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7bd30d2f-f0d0-4fb0-8a71-1d82af263cdc · outbound
NoLoCo: No-all-reduce Low Communication Training Method for Large Models LLaMA: Open and Efficient Foundation Language Models
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation acc50d72-f732-4238-9b91-c6dec2b411bf · outbound
NoLoCo: No-all-reduce Low Communication Training Method for Large Models Understanding Short-Horizon Bias in Stochastic Meta-Optimization
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 220f8bba-cf8e-446a-b0d3-556a1f5f6b26 · outbound
NoLoCo: No-all-reduce Low Communication Training Method for Large Models LLaVA-Mini: Efficient Image and Video Large Multimodal Models with One Vision Token
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 690eda80-e49a-436d-882c-c9a5acbceded · outbound
NoLoCo: No-all-reduce Low Communication Training Method for Large Models OPT: Open Pre-trained Transformer Language Models
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f07824be-df69-48a6-93ae-a5650d60d4c5 · outbound
NoLoCo: No-all-reduce Low Communication Training Method for Large Models MiniGPT-4: Enhancing Vision-Language Understanding with Advanced Large Language Models
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 53121814-be8b-4886-b40d-58a300d2c5a5 · outbound
NoLoCo: No-all-reduce Low Communication Training Method for Large Models Megatron-LM: Training Multi-Billion Parameter Language Models Using Model Parallelism
Reference 2013
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 98f1ddbe-3a72-4627-8ce1-11e6c3bbf1d3 · outbound
NoLoCo: No-all-reduce Low Communication Training Method for Large Models Qwen2.5-Omni Technical Report
Reference 2018
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 447711e1-fe9b-42e6-9541-dab3815cd531 · outbound
NoLoCo: No-all-reduce Low Communication Training Method for Large Models Tree Attention: Topology-aware Decoding for Long-Context Attention on GPU clusters
Reference 2019
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 55ee83f6-9f03-4f0c-a4d7-a8b5b4219cc9 · outbound
NoLoCo: No-all-reduce Low Communication Training Method for Large Models Boosting Asynchronous Decentralized Learning with Model Fragmentation
Reference 2020
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 407c224f-5d82-43e0-9a22-a0ba775b36cf · outbound
NoLoCo: No-all-reduce Low Communication Training Method for Large Models DiLoCo: Distributed Low-Communication Training of Language Models
Reference 2021
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1faccbb5-2426-49f5-9ba6-a16d4417fef3 · outbound
NoLoCo: No-all-reduce Low Communication Training Method for Large Models Video-LLaVA: Learning United Visual Representation by Alignment Before Projection
Reference 2022
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4d229f63-2d0b-43a6-bd9c-1b00d23b3a15 · outbound
NoLoCo: No-all-reduce Low Communication Training Method for Large Models DiPaCo: Distributed Path Composition
Reference 2023
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 280f9081-1285-4bea-b77f-4aa7a991a3ec · outbound
NoLoCo: No-all-reduce Low Communication Training Method for Large Models A Survey on Mixture of Experts in Large Language Models
Reference 2024
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a13ce9fc-8aa6-4add-9d92-7032a549a2ea · outbound
NoLoCo: No-all-reduce Low Communication Training Method for Large Models Documenting Large Webtext Corpora: A Case Study on the Colossal Clean Crawled Corpus
Reference 2025
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9f4ddb48-f020-4438-8ec9-11deb9df7c3c · inbound
On the Surprising Effectiveness of a Single Global Merging in Decentralized Learning NoLoCo: No-all-reduce Low Communication Training Method for Large Models
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation d5a3b6dc-44df-48ab-b0e6-d5b776525c35 · inbound
Decoupled DiLoCo for Resilient Distributed Pre-training NoLoCo: No-all-reduce Low Communication Training Method for Large Models
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 1369cf40-bbcb-47e1-8b3d-2acc0ba81afd · inbound
HeLoCo: Efficient asynchronous low-communication training under data and device heterogeneity NoLoCo: No-all-reduce Low Communication Training Method for Large Models
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 4ee2c74a-bab3-43c8-8200-9d4ed470519e · inbound
Unifying Local Communications and Local Updates for LLM Pretraining NoLoCo: No-all-reduce Low Communication Training Method for Large Models
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 20fcc947-712b-4d19-80f8-a8be10b42200 · inbound
Not Every Sync Is Safe: Calibrated DiLoCo Scheduling for Shared AI Infrastructure NoLoCo: No-all-reduce Low Communication Training Method for Large Models
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.