Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-08T10:20:02.783216Z
Paper Citation Record · LEDGER
As of 8 August 2026, this Paper Citation Record lists 49 of 49 outbound references and 0 inbound Pith citation observations for arXiv:2502.08145.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-08T10:20:02.783216Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
49 of 49 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation cfe60ddb-b8d2-4770-b176-54ab30e9e410 · outbound
Democratizing AI: Open-source Scalable LLM Training on GPU-based Supercomputers Super: Sub-graph parallelism for transformers,
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation e6e52360-d80b-4c1f-a0da-591463fe3c47 · outbound
Democratizing AI: Open-source Scalable LLM Training on GPU-based Supercomputers Scaling distributed deep learning work- loads beyond the memory capacity with karma,
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 173818ab-a30e-46b0-8e04-3a0b664a992f · outbound
Democratizing AI: Open-source Scalable LLM Training on GPU-based Supercomputers Forge: Pre-training open foundation models for science,
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 05f24982-92c5-4040-a46f-8d80de8a20a7 · outbound
Democratizing AI: Open-source Scalable LLM Training on GPU-based Supercomputers Optimizing distributed training on frontier for large language models,
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation ae7568c9-1fd7-465d-8c55-ee1c7922a044 · outbound
Democratizing AI: Open-source Scalable LLM Training on GPU-based Supercomputers Using deepspeed and megatron to train megatron-turing nlg 530b, a large-scale generative language model,
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation fb0670d6-8efc-4d2e-b247-9226f418704c · outbound
Democratizing AI: Open-source Scalable LLM Training on GPU-based Supercomputers Efficient Large-Scale Language Model Training on GPU Clusters Using Megatron-LM
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e5e73c07-1c20-43a9-b9c3-b2cccfb97aa8 · outbound
Democratizing AI: Open-source Scalable LLM Training on GPU-based Supercomputers MegaScale: Scaling large language model training to more than 10,000 GPUs,
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 5fb685db-f7b6-4a83-b6f5-0fb09d67e5db · outbound
Democratizing AI: Open-source Scalable LLM Training on GPU-based Supercomputers Google cloud demonstrates the world’s largest distributed training job for large language models across 50000+ tpu v5e chips,
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation d31439de-b84f-43b5-ab93-7aff0e022af2 · outbound
Democratizing AI: Open-source Scalable LLM Training on GPU-based Supercomputers AxoNN: An asynchronous, message-driven parallel framework for extreme-scale deep learning,
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 34d2b012-dcf0-4de7-9aac-58efaec82c63 · outbound
Democratizing AI: Open-source Scalable LLM Training on GPU-based Supercomputers Exploiting sparsity in pruned neural networks to optimize large model training,
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 92cc03c1-322f-49ff-9c7e-96cfae54dcd9 · outbound
Democratizing AI: Open-source Scalable LLM Training on GPU-based Supercomputers Zero: Memory optimizations toward training trillion parameter models,
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 15d514db-84c9-449c-8c4d-b2e5864ee6fd · outbound
Democratizing AI: Open-source Scalable LLM Training on GPU-based Supercomputers Pytorch fsdp: Experiences on scaling fully sharded data parallel,
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 3adfc4b9-0a73-4fa5-ad42-26269b8f8431 · outbound
Democratizing AI: Open-source Scalable LLM Training on GPU-based Supercomputers Megatron-lm: Training multi-billion parameter language models using model parallelism,
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation a4b70986-2b49-46e6-8dc1-4fe17fee7aca · outbound
Democratizing AI: Open-source Scalable LLM Training on GPU-based Supercomputers GPipe: efficient training of giant neural networks using pipeline parallelism,
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation d1a34e15-d91d-4c5d-9bd6-52060981d60c · outbound
Democratizing AI: Open-source Scalable LLM Training on GPU-based Supercomputers Deepspeed: Extreme-scale model training for everyone,
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation e8091f37-6dfe-4fa1-8251-e0e7012908a5 · outbound
Democratizing AI: Open-source Scalable LLM Training on GPU-based Supercomputers A hybrid tensor-expert-data parallelism approach to optimize mixture-of-experts training,
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 1dc7ce9b-2044-4557-8c17-4858138b72d0 · outbound
Democratizing AI: Open-source Scalable LLM Training on GPU-based Supercomputers GPT-NeoX: Large Scale Autoregressive Language Modeling in PyTorch,
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 2a4c607d-2d59-4ad5-aba4-f0ec474863d6 · outbound
Democratizing AI: Open-source Scalable LLM Training on GPU-based Supercomputers Alpa: Automating Inter- and Intra-Operator Parallelism for Distributed Deep Learning
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6f350aae-1cf5-48ed-b4fd-71628a22086c · outbound
Democratizing AI: Open-source Scalable LLM Training on GPU-based Supercomputers Colossal-AI: a unified deep learning system for large-scale parallel training,
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 8854f878-52af-44af-9404-6ff66b5a0ad9 · outbound
Democratizing AI: Open-source Scalable LLM Training on GPU-based Supercomputers Llama 2: Open foundation and fine-tuned chat models,
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 8b3f9544-2386-40d4-9509-223cf5ea0c6e · outbound
Democratizing AI: Open-source Scalable LLM Training on GPU-based Supercomputers Exploring the Limits of Transfer Learning with a Unified Text-to-Text Transformer
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9ba13cb0-3cdc-4355-a15a-12751caef3c7 · outbound
Democratizing AI: Open-source Scalable LLM Training on GPU-based Supercomputers LBANN: livermore big artificial neural network HPC toolkit,
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation bfd7db8b-9d6c-4dd3-bd2e-e40fcaf5ce0d · outbound
Democratizing AI: Open-source Scalable LLM Training on GPU-based Supercomputers Nvidia selene supercomputer,
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation d6c41d50-7547-4753-9b5c-04cfec9da0fd · outbound
Democratizing AI: Open-source Scalable LLM Training on GPU-based Supercomputers Frontier: Exploring exascale,
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 02f325f9-583f-4c53-9893-f82b4e765e40 · outbound
Democratizing AI: Open-source Scalable LLM Training on GPU-based Supercomputers A three-dimensional approach to parallel matrix multiplication,
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 38b79e32-95e6-49cb-a25c-afe97b3a2065 · outbound
Democratizing AI: Open-source Scalable LLM Training on GPU-based Supercomputers ZeRO++: Extremely efficient collective communication for large model training,
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation fb5c813e-084a-4de4-9796-71e62a58cd82 · outbound
Democratizing AI: Open-source Scalable LLM Training on GPU-based Supercomputers Improving the performance of collective operations in mpich,
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 961544f7-2ae5-40f2-877b-6060c90bdeb0 · outbound
Democratizing AI: Open-source Scalable LLM Training on GPU-based Supercomputers Optimization of collective reduction operations,
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 027f16fe-75f9-4c37-b9be-d0606c3646e4 · outbound
Democratizing AI: Open-source Scalable LLM Training on GPU-based Supercomputers Improving communication performance in dense linear algebra via topology aware collectives,
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 2856382c-b8dc-419c-9ece-b2bb727d5dbb · outbound
Democratizing AI: Open-source Scalable LLM Training on GPU-based Supercomputers Mapping applications with collectives over sub-communicators on torus networks,
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation ffae8d80-2e29-432e-8d87-6220294e365e · outbound
Democratizing AI: Open-source Scalable LLM Training on GPU-based Supercomputers RAHTM: Routing- algorithm aware hierarchical task mapping,
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 18f394f7-eb24-445c-80cb-9b33769fb30e · outbound
Democratizing AI: Open-source Scalable LLM Training on GPU-based Supercomputers Optimizing the performance of parallel applications on a 5D torus via task mapping,
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation fa8248a8-893e-42fb-8106-6fa3d1da566c · outbound
Democratizing AI: Open-source Scalable LLM Training on GPU-based Supercomputers Supervised learning based algorithm selection for deep neural networks,
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation bcbddde1-a435-4540-98b6-c815a313d1fd · outbound
Democratizing AI: Open-source Scalable LLM Training on GPU-based Supercomputers Language Models are Few-Shot Learners
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 25f95a0b-c2c5-408b-85ae-6bb5b9478acc · outbound
Democratizing AI: Open-source Scalable LLM Training on GPU-based Supercomputers Attention Is All You Need
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b6d68b7c-b16a-48b5-9784-ea7c3494d9d0 · outbound
Democratizing AI: Open-source Scalable LLM Training on GPU-based Supercomputers Bigscience large open-science open-access multilingual language model,
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 9a5e2c51-0a27-4a59-85c6-11fea84bdb77 · outbound
Democratizing AI: Open-source Scalable LLM Training on GPU-based Supercomputers Language models are unsupervised multitask learners,
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 3a6163ec-71d9-4a64-9421-e53764fd4d5f · outbound
Democratizing AI: Open-source Scalable LLM Training on GPU-based Supercomputers Training Deep Nets with Sublinear Memory Cost
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 933cbc24-0067-4915-9828-d982de5f4ef6 · outbound
Democratizing AI: Open-source Scalable LLM Training on GPU-based Supercomputers A Study of BFLOAT16 for Deep Learning Training
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 43ef57a5-c4a9-4925-8a90-b4a201bfa18f · outbound
Democratizing AI: Open-source Scalable LLM Training on GPU-based Supercomputers Unresolved cited work
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 7920df23-0bb7-4aa8-a738-cf3dca58ebf1 · outbound
Democratizing AI: Open-source Scalable LLM Training on GPU-based Supercomputers Interactive investigation of traffic congestion on fat-tree networks using TreeScope,
Reference 42
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 0f67b059-0295-489a-8c68-5dbbec1579ef · outbound
Democratizing AI: Open-source Scalable LLM Training on GPU-based Supercomputers Quantifying I/O and communication traffic interference on dragonfly networks equipped with burst buffers,
Reference 43
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 26feff98-4dfa-4171-9546-48f55f5f623d · outbound
Democratizing AI: Open-source Scalable LLM Training on GPU-based Supercomputers Quantifying memorization across neural language models,
Reference 44
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 648bc8e9-bf90-4fe9-b9dd-27e200a2238b · outbound
Democratizing AI: Open-source Scalable LLM Training on GPU-based Supercomputers The times sues openai and microsoft over ai use of copyrighted work,
Reference 45
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 9efce775-2041-4e4a-ab0f-cbf5d21b4bf6 · outbound
Democratizing AI: Open-source Scalable LLM Training on GPU-based Supercomputers Extracting training data from large language models,
Reference 46
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9d69f755-c00b-4b91-97d0-6c7080ca91fa · outbound
Democratizing AI: Open-source Scalable LLM Training on GPU-based Supercomputers Pythia: A suite for analyzing large language models across training and scaling,
Reference 47
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 178cdb6d-caa0-4152-ae75-80a4004fa1dc · outbound
Democratizing AI: Open-source Scalable LLM Training on GPU-based Supercomputers Tinyllama: An open-source small language model,
Reference 48
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation df901eb7-3427-494b-baae-7e033835fb2a · outbound
Democratizing AI: Open-source Scalable LLM Training on GPU-based Supercomputers The llama 3 herd of models,
Reference 49
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 2145c077-2e84-45e2-9e97-4ed4002bfff4 · outbound
Democratizing AI: Open-source Scalable LLM Training on GPU-based Supercomputers Be like a Goldfish, Don't Memorize! Mitigating Memorization in Generative LLMs
Reference 50
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
No inbound Pith citation observations are available.