Pith. sign in

Paper Citation Record · LEDGER

Democratizing AI: Open-source Scalable LLM Training on GPU-based Supercomputers

As of 8 August 2026, this Paper Citation Record lists 49 of 49 outbound references and 0 inbound Pith citation observations for arXiv:2502.08145.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2502.08145 v1

Coverage vector

measured 49 of 49 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-08T10:20:02.783216Z

measured 49 of 49 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

49 of 49 outbound references displayed

  • verified exact5
  • verified fuzzy33
  • unresolved10
  • parse uncertain0
  • malformed identifier1
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation cfe60ddb-b8d2-4770-b176-54ab30e9e410 · outbound

This paper cites Super: Sub-graph parallelism for transformers,.

Democratizing AI: Open-source Scalable LLM Training on GPU-based Supercomputers Super: Sub-graph parallelism for transformers,

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T10:20:05.277436Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-08T10:20:02.624115Z digest=sha256:8b21b5467cc45da074c7f78cbe0889e0e625bee089a4f197b87616bbfe5be5d1

Observation e6e52360-d80b-4c1f-a0da-591463fe3c47 · outbound

This paper cites Scaling distributed deep learning work- loads beyond the memory capacity with karma,.

Democratizing AI: Open-source Scalable LLM Training on GPU-based Supercomputers Scaling distributed deep learning work- loads beyond the memory capacity with karma,

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T10:20:05.268993Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-08T10:20:02.628041Z digest=sha256:c1d74b86ec14b0b86ed5c6f52b327f36b1015e9c91126cf4bdfb7159b5d361de

Observation 173818ab-a30e-46b0-8e04-3a0b664a992f · outbound

This paper cites Forge: Pre-training open foundation models for science,.

Democratizing AI: Open-source Scalable LLM Training on GPU-based Supercomputers Forge: Pre-training open foundation models for science,

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T10:20:05.261411Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-08T10:20:02.631258Z digest=sha256:537e0bd86d7f8d15f7a8564d18798430082fc47c00d1fbf4ed8e8a1e17065574

Observation 05f24982-92c5-4040-a46f-8d80de8a20a7 · outbound

This paper cites Optimizing distributed training on frontier for large language models,.

Democratizing AI: Open-source Scalable LLM Training on GPU-based Supercomputers Optimizing distributed training on frontier for large language models,

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T10:20:05.251706Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-08T10:20:02.634237Z digest=sha256:987bf83693b2038975dbe59ebccd0551501bf49a49eaad68ea800d3bafbc1d79

Observation ae7568c9-1fd7-465d-8c55-ee1c7922a044 · outbound

This paper cites Using deepspeed and megatron to train megatron-turing nlg 530b, a large-scale generative language model,.

Democratizing AI: Open-source Scalable LLM Training on GPU-based Supercomputers Using deepspeed and megatron to train megatron-turing nlg 530b, a large-scale generative language model,

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T10:20:05.242863Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-08T10:20:02.638603Z digest=sha256:879be5150c6e747a86df4fc848e5fc4852906abc739342291fb461e3b303098e

Observation fb0670d6-8efc-4d2e-b247-9226f418704c · outbound

This paper cites Efficient Large-Scale Language Model Training on GPU Clusters Using Megatron-LM.

Democratizing AI: Open-source Scalable LLM Training on GPU-based Supercomputers Efficient Large-Scale Language Model Training on GPU Clusters Using Megatron-LM

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-08T10:20:02.642128Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T10:20:02.642128Z digest=sha256:a6200d1aee183845070a555fb3460e1a90f117073439c119362ee9652ea90e71

Observation e5e73c07-1c20-43a9-b9c3-b2cccfb97aa8 · outbound

This paper cites MegaScale: Scaling large language model training to more than 10,000 GPUs,.

Democratizing AI: Open-source Scalable LLM Training on GPU-based Supercomputers MegaScale: Scaling large language model training to more than 10,000 GPUs,

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T10:20:05.234090Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-08T10:20:02.645980Z digest=sha256:413bf175ed3346c87a8e7f8f9c9340f214a1282c0098b0c1a8c0d5d5af2d16d6

Observation 5fb685db-f7b6-4a83-b6f5-0fb09d67e5db · outbound

This paper cites Google cloud demonstrates the world’s largest distributed training job for large language models across 50000+ tpu v5e chips,.

Democratizing AI: Open-source Scalable LLM Training on GPU-based Supercomputers Google cloud demonstrates the world’s largest distributed training job for large language models across 50000+ tpu v5e chips,

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T10:20:05.224647Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-08T10:20:02.649415Z digest=sha256:8a17fd93c0b125d9cd4acc93f679ba4619dd2bec2fc3f0a3cd5386808f118c25

Observation d31439de-b84f-43b5-ab93-7aff0e022af2 · outbound

This paper cites AxoNN: An asynchronous, message-driven parallel framework for extreme-scale deep learning,.

Democratizing AI: Open-source Scalable LLM Training on GPU-based Supercomputers AxoNN: An asynchronous, message-driven parallel framework for extreme-scale deep learning,

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T10:20:05.215715Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-08T10:20:02.652845Z digest=sha256:f0f99e1ae0d80524170cf8b3317ded3449da1b17db1bbb98dc3ae505535c58f2

Observation 34d2b012-dcf0-4de7-9aac-58efaec82c63 · outbound

This paper cites Exploiting sparsity in pruned neural networks to optimize large model training,.

Democratizing AI: Open-source Scalable LLM Training on GPU-based Supercomputers Exploiting sparsity in pruned neural networks to optimize large model training,

Reference 10

Resolution
verified exact
arxiv_id_nonexistent, observed 2026-08-08T10:20:04.956493Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-08T10:20:02.655642Z digest=sha256:ef8888580335f1fc9fc5ac0378b6947a51a45687f38f359f0842f87029baf94a

Observation 92cc03c1-322f-49ff-9c7e-96cfae54dcd9 · outbound

This paper cites Zero: Memory optimizations toward training trillion parameter models,.

Democratizing AI: Open-source Scalable LLM Training on GPU-based Supercomputers Zero: Memory optimizations toward training trillion parameter models,

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T10:20:05.206332Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-08T10:20:02.659372Z digest=sha256:4914a2974585538af735172053a4e36ea86c14cb513db51ca5a107d4d5316b50

Observation 15d514db-84c9-449c-8c4d-b2e5864ee6fd · outbound

This paper cites Pytorch fsdp: Experiences on scaling fully sharded data parallel,.

Democratizing AI: Open-source Scalable LLM Training on GPU-based Supercomputers Pytorch fsdp: Experiences on scaling fully sharded data parallel,

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T10:20:05.197454Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-08T10:20:02.662349Z digest=sha256:408f7c3e269cf6974c6a737e4b17a488fa4070fb1a7330d8dc0976bf111200ca

Observation 3adfc4b9-0a73-4fa5-ad42-26269b8f8431 · outbound

This paper cites Megatron-lm: Training multi-billion parameter language models using model parallelism,.

Democratizing AI: Open-source Scalable LLM Training on GPU-based Supercomputers Megatron-lm: Training multi-billion parameter language models using model parallelism,

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T10:20:05.179145Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-08T10:20:02.668560Z digest=sha256:22048e7036935baab28033fc8700397d0dad1c61c6367047f05354a5998a7267

Observation a4b70986-2b49-46e6-8dc1-4fe17fee7aca · outbound

This paper cites GPipe: efficient training of giant neural networks using pipeline parallelism,.

Democratizing AI: Open-source Scalable LLM Training on GPU-based Supercomputers GPipe: efficient training of giant neural networks using pipeline parallelism,

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T10:20:05.170498Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-08T10:20:02.671345Z digest=sha256:21722001bdca679227488eb77658e28769363877bebfe90217105a43f7198b71

Observation d1a34e15-d91d-4c5d-9bd6-52060981d60c · outbound

This paper cites Deepspeed: Extreme-scale model training for everyone,.

Democratizing AI: Open-source Scalable LLM Training on GPU-based Supercomputers Deepspeed: Extreme-scale model training for everyone,

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T10:20:05.162555Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-08T10:20:02.674012Z digest=sha256:808d574c3e10b7684f1ff7f0de124a579f4edcca146bc38e41a8f5e4999a12e5

Observation e8091f37-6dfe-4fa1-8251-e0e7012908a5 · outbound

This paper cites A hybrid tensor-expert-data parallelism approach to optimize mixture-of-experts training,.

Democratizing AI: Open-source Scalable LLM Training on GPU-based Supercomputers A hybrid tensor-expert-data parallelism approach to optimize mixture-of-experts training,

Reference 17

Resolution
verified exact
arxiv_id_nonexistent, observed 2026-08-08T10:20:04.609477Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-08T10:20:02.677289Z digest=sha256:29e06e339a9d0786f69b7af5a8edb315a083f710c376b83fd2e67d38c3b1bc2f

Observation 1dc7ce9b-2044-4557-8c17-4858138b72d0 · outbound

This paper cites GPT-NeoX: Large Scale Autoregressive Language Modeling in PyTorch,.

Democratizing AI: Open-source Scalable LLM Training on GPU-based Supercomputers GPT-NeoX: Large Scale Autoregressive Language Modeling in PyTorch,

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T10:20:05.152654Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-08T10:20:02.680393Z digest=sha256:297503cf534a9e56d30624e7ef81db8109a391ebcdc084c25b1b1720caab16d3

Observation 2a4c607d-2d59-4ad5-aba4-f0ec474863d6 · outbound

This paper cites Alpa: Automating Inter- and Intra-Operator Parallelism for Distributed Deep Learning.

Democratizing AI: Open-source Scalable LLM Training on GPU-based Supercomputers Alpa: Automating Inter- and Intra-Operator Parallelism for Distributed Deep Learning

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-08T10:20:02.683674Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T10:20:02.683674Z digest=sha256:62ea52c625b6bcb6945c28f336b4edac0586f636b8a99e736dc224c47308eecd

Observation 6f350aae-1cf5-48ed-b4fd-71628a22086c · outbound

This paper cites Colossal-AI: a unified deep learning system for large-scale parallel training,.

Democratizing AI: Open-source Scalable LLM Training on GPU-based Supercomputers Colossal-AI: a unified deep learning system for large-scale parallel training,

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T10:20:05.144149Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-08T10:20:02.686844Z digest=sha256:09fca981bc4a32910cfee132398e191687f844bac027345c7d46af2188c2b323

Observation 8854f878-52af-44af-9404-6ff66b5a0ad9 · outbound

This paper cites Llama 2: Open foundation and fine-tuned chat models,.

Democratizing AI: Open-source Scalable LLM Training on GPU-based Supercomputers Llama 2: Open foundation and fine-tuned chat models,

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T10:20:05.133794Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-08T10:20:02.689764Z digest=sha256:af118bbefe7fc5057db734bda9ed5895ef2b210191d116bc1467e16e5dbca967

Observation 8b3f9544-2386-40d4-9509-223cf5ea0c6e · outbound

This paper cites Exploring the Limits of Transfer Learning with a Unified Text-to-Text Transformer.

Democratizing AI: Open-source Scalable LLM Training on GPU-based Supercomputers Exploring the Limits of Transfer Learning with a Unified Text-to-Text Transformer

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-08T10:20:02.692885Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T10:20:02.692885Z digest=sha256:5072ee35775a91d17e561b479c3d7af1a64a3fcf3a0fd477d6cfdfd09e83d060

Observation 9ba13cb0-3cdc-4355-a15a-12751caef3c7 · outbound

This paper cites LBANN: livermore big artificial neural network HPC toolkit,.

Democratizing AI: Open-source Scalable LLM Training on GPU-based Supercomputers LBANN: livermore big artificial neural network HPC toolkit,

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T10:20:05.125390Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-08T10:20:02.696500Z digest=sha256:690be92af7405bb218109127df6e34cb7ace2afac6d8fa0690f335e128e6c1eb

Observation bfd7db8b-9d6c-4dd3-bd2e-e40fcaf5ce0d · outbound

This paper cites Nvidia selene supercomputer,.

Democratizing AI: Open-source Scalable LLM Training on GPU-based Supercomputers Nvidia selene supercomputer,

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T10:20:05.117015Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-08T10:20:02.699888Z digest=sha256:aaf155a25ab2da7cfb4ec51e1d899a77baf09672b34078c78008326bd2e4d346

Observation d6c41d50-7547-4753-9b5c-04cfec9da0fd · outbound

This paper cites Frontier: Exploring exascale,.

Democratizing AI: Open-source Scalable LLM Training on GPU-based Supercomputers Frontier: Exploring exascale,

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T10:20:05.108866Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-08T10:20:02.703399Z digest=sha256:be90b0a3be28a814e41d1115bd68e20bb2189a0b961239ff18eb3f11682ab5f1

Observation 02f325f9-583f-4c53-9893-f82b4e765e40 · outbound

This paper cites A three-dimensional approach to parallel matrix multiplication,.

Democratizing AI: Open-source Scalable LLM Training on GPU-based Supercomputers A three-dimensional approach to parallel matrix multiplication,

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T10:20:05.100532Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-08T10:20:02.706230Z digest=sha256:734d022aa29ab232ee97dd24a8f619dd6ed9024739a056ff9f6faedc8426e871

Observation 38b79e32-95e6-49cb-a25c-afe97b3a2065 · outbound

This paper cites ZeRO++: Extremely efficient collective communication for large model training,.

Democratizing AI: Open-source Scalable LLM Training on GPU-based Supercomputers ZeRO++: Extremely efficient collective communication for large model training,

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T10:20:05.188445Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-08T10:20:02.709021Z digest=sha256:1e01e96b74b20f5252c39ffa7e12d46c1865c4f975ca81a694c797c5fe73f0ef

Observation fb5c813e-084a-4de4-9796-71e62a58cd82 · outbound

This paper cites Improving the performance of collective operations in mpich,.

Democratizing AI: Open-source Scalable LLM Training on GPU-based Supercomputers Improving the performance of collective operations in mpich,

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T10:20:05.091221Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-08T10:20:02.712838Z digest=sha256:0af127c51fe147a9419605eb31de1117d66611117d6862ad249211ca15aed2f1

Observation 961544f7-2ae5-40f2-877b-6060c90bdeb0 · outbound

This paper cites Optimization of collective reduction operations,.

Democratizing AI: Open-source Scalable LLM Training on GPU-based Supercomputers Optimization of collective reduction operations,

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T10:20:05.082585Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-08T10:20:02.715458Z digest=sha256:bfc2bacc9994160c24fd38e2cfd8501fab6f2a2124473bf66cdeb632d46d3ba3

Observation 027f16fe-75f9-4c37-b9be-d0606c3646e4 · outbound

This paper cites Improving communication performance in dense linear algebra via topology aware collectives,.

Democratizing AI: Open-source Scalable LLM Training on GPU-based Supercomputers Improving communication performance in dense linear algebra via topology aware collectives,

Reference 30

Resolution
verified exact
arxiv_id_nonexistent, observed 2026-08-08T10:20:03.301128Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-08T10:20:02.718427Z digest=sha256:1e0441668b17ce4246a52895fde11a917a726f217e64e5be55076038e6d4cd86

Observation 2856382c-b8dc-419c-9ece-b2bb727d5dbb · outbound

This paper cites Mapping applications with collectives over sub-communicators on torus networks,.

Democratizing AI: Open-source Scalable LLM Training on GPU-based Supercomputers Mapping applications with collectives over sub-communicators on torus networks,

Reference 31

Resolution
malformed identifier
doi_truncated, observed 2026-08-08T10:20:02.820258Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-08T10:20:02.721918Z digest=sha256:da246453a4623421622cb85c0479a2016a07b3422c421ff87f94a2cbbfbb2a3f

Observation ffae8d80-2e29-432e-8d87-6220294e365e · outbound

This paper cites RAHTM: Routing- algorithm aware hierarchical task mapping,.

Democratizing AI: Open-source Scalable LLM Training on GPU-based Supercomputers RAHTM: Routing- algorithm aware hierarchical task mapping,

Reference 32

Resolution
verified exact
doi, observed 2026-08-08T10:20:02.809304Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-08T10:20:02.725877Z digest=sha256:102af208158123a8041cb7425b5a4ca0bdbd9b9daba373aee36f064e226ed7e3

Observation 18f394f7-eb24-445c-80cb-9b33769fb30e · outbound

This paper cites Optimizing the performance of parallel applications on a 5D torus via task mapping,.

Democratizing AI: Open-source Scalable LLM Training on GPU-based Supercomputers Optimizing the performance of parallel applications on a 5D torus via task mapping,

Reference 33

Resolution
verified exact
arxiv_id_nonexistent, observed 2026-08-08T10:20:03.048400Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-08T10:20:02.728858Z digest=sha256:cdd93b7b897dbe9dc963632037f755aea80304678c6b6edd10b7e9a42cf2948c

Observation fa8248a8-893e-42fb-8106-6fa3d1da566c · outbound

This paper cites Supervised learning based algorithm selection for deep neural networks,.

Democratizing AI: Open-source Scalable LLM Training on GPU-based Supercomputers Supervised learning based algorithm selection for deep neural networks,

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T10:20:05.073494Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-08T10:20:02.732776Z digest=sha256:6ee956858949b988792990873bf78f8ba750c5d442b83e767bb04e98882119dc

Observation bcbddde1-a435-4540-98b6-c815a313d1fd · outbound

This paper cites Language Models are Few-Shot Learners.

Democratizing AI: Open-source Scalable LLM Training on GPU-based Supercomputers Language Models are Few-Shot Learners

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-08T10:20:02.735880Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T10:20:02.735880Z digest=sha256:a312a340ffbbf5e4088cc015153a02cb740dd3aa1fb80f34ce0ec1cee7d4280c

Observation 25f95a0b-c2c5-408b-85ae-6bb5b9478acc · outbound

This paper cites Attention Is All You Need.

Democratizing AI: Open-source Scalable LLM Training on GPU-based Supercomputers Attention Is All You Need

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-08T10:20:02.739697Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T10:20:02.739697Z digest=sha256:cb2cb2c4b32f56637c10de40d3288666966b0cc8c830fe56de5de8c7adda7950

Observation b6d68b7c-b16a-48b5-9784-ea7c3494d9d0 · outbound

This paper cites Bigscience large open-science open-access multilingual language model,.

Democratizing AI: Open-source Scalable LLM Training on GPU-based Supercomputers Bigscience large open-science open-access multilingual language model,

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T10:20:05.065419Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-08T10:20:02.742736Z digest=sha256:1c45691a2c9016775b6d0663f2a2f49d2798386976ad5dc16f1a750997951c44

Observation 9a5e2c51-0a27-4a59-85c6-11fea84bdb77 · outbound

This paper cites Language models are unsupervised multitask learners,.

Democratizing AI: Open-source Scalable LLM Training on GPU-based Supercomputers Language models are unsupervised multitask learners,

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T10:20:05.056115Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-08T10:20:02.746066Z digest=sha256:a417821051a5e91f364b349787bb9a6cba3fc8989dc8d6e111383e9774c10f1e

Observation 3a6163ec-71d9-4a64-9421-e53764fd4d5f · outbound

This paper cites Training Deep Nets with Sublinear Memory Cost.

Democratizing AI: Open-source Scalable LLM Training on GPU-based Supercomputers Training Deep Nets with Sublinear Memory Cost

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-08T10:20:02.748714Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T10:20:02.748714Z digest=sha256:110bab1bd8b15fd98dfacff9eaea4e49ea10e9bfd3792cf26975912ab99e2f6d

Observation 933cbc24-0067-4915-9828-d982de5f4ef6 · outbound

This paper cites A Study of BFLOAT16 for Deep Learning Training.

Democratizing AI: Open-source Scalable LLM Training on GPU-based Supercomputers A Study of BFLOAT16 for Deep Learning Training

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-08T10:20:02.751774Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T10:20:02.751774Z digest=sha256:3277afd51011a3e2cc0d87f766e22259ef72cb8f85070cf4f1fefcec50c56bec

Observation 43ef57a5-c4a9-4925-8a90-b4a201bfa18f · outbound

This paper cites an unresolved cited work.

Democratizing AI: Open-source Scalable LLM Training on GPU-based Supercomputers Unresolved cited work

Reference 41

Resolution
unresolved
raw_fallback, observed 2026-08-08T10:20:05.046128Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-08T10:20:02.754450Z digest=sha256:f1457fc786d5be4e43cda6aa33736600b08a399daf0af7d3bac440ffed6c166b

Observation 7920df23-0bb7-4aa8-a738-cf3dca58ebf1 · outbound

This paper cites Interactive investigation of traffic congestion on fat-tree networks using TreeScope,.

Democratizing AI: Open-source Scalable LLM Training on GPU-based Supercomputers Interactive investigation of traffic congestion on fat-tree networks using TreeScope,

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T10:20:05.036954Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-08T10:20:02.757875Z digest=sha256:d42065629ef8ccd82a8fc11f552f3e89f66940390e3ec15cbed0c18a1d819665

Observation 0f67b059-0295-489a-8c68-5dbbec1579ef · outbound

This paper cites Quantifying I/O and communication traffic interference on dragonfly networks equipped with burst buffers,.

Democratizing AI: Open-source Scalable LLM Training on GPU-based Supercomputers Quantifying I/O and communication traffic interference on dragonfly networks equipped with burst buffers,

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T10:20:05.028284Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-08T10:20:02.761481Z digest=sha256:7ba4322e2fa3ef9a402823ef1f596c715b6d4292cbef8afdfe0ce7dc0e58d2ff

Observation 26feff98-4dfa-4171-9546-48f55f5f623d · outbound

This paper cites Quantifying memorization across neural language models,.

Democratizing AI: Open-source Scalable LLM Training on GPU-based Supercomputers Quantifying memorization across neural language models,

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T10:20:05.018885Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-08T10:20:02.764626Z digest=sha256:6a6bd07def2a847ff7ad7c59543c7016126bd5e1c8d5ab64751c3f6c945178e0

Observation 648bc8e9-bf90-4fe9-b9dd-27e200a2238b · outbound

This paper cites The times sues openai and microsoft over ai use of copyrighted work,.

Democratizing AI: Open-source Scalable LLM Training on GPU-based Supercomputers The times sues openai and microsoft over ai use of copyrighted work,

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T10:20:05.008971Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-08T10:20:02.768307Z digest=sha256:ed0225b494d4c5336d231a52c6e4bfc9f364cb443b5cf6ea6151c4c288108ba6

Observation 9efce775-2041-4e4a-ab0f-cbf5d21b4bf6 · outbound

This paper cites Extracting training data from large language models,.

Democratizing AI: Open-source Scalable LLM Training on GPU-based Supercomputers Extracting training data from large language models,

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-08T10:20:02.771544Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T10:20:02.771544Z digest=sha256:f0ca02c22494cdf14c556c469324d27f58df63f3b8eddf03056d043f2f5fe46b

Observation 9d69f755-c00b-4b91-97d0-6c7080ca91fa · outbound

This paper cites Pythia: A suite for analyzing large language models across training and scaling,.

Democratizing AI: Open-source Scalable LLM Training on GPU-based Supercomputers Pythia: A suite for analyzing large language models across training and scaling,

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T10:20:04.994313Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-08T10:20:02.774029Z digest=sha256:280569579b01381000e620a4d14a79b1f8a3b4107849b025df4d93116e679f2b

Observation 178cdb6d-caa0-4152-ae75-80a4004fa1dc · outbound

This paper cites Tinyllama: An open-source small language model,.

Democratizing AI: Open-source Scalable LLM Training on GPU-based Supercomputers Tinyllama: An open-source small language model,

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T10:20:04.985533Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-08T10:20:02.777191Z digest=sha256:7bb117b9b61cdc0c991b6cb07553fac7a3ae3a2a6897273f86f27af2f47494da

Observation df901eb7-3427-494b-baae-7e033835fb2a · outbound

This paper cites The llama 3 herd of models,.

Democratizing AI: Open-source Scalable LLM Training on GPU-based Supercomputers The llama 3 herd of models,

Reference 49

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T10:20:04.976637Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-08T10:20:02.780659Z digest=sha256:574057eb0c98ec08aa877ec867c25a11056b74aac062339f92bec944d74b20ca

Observation 2145c077-2e84-45e2-9e97-4ed4002bfff4 · outbound

This paper cites Be like a Goldfish, Don't Memorize! Mitigating Memorization in Generative LLMs.

Democratizing AI: Open-source Scalable LLM Training on GPU-based Supercomputers Be like a Goldfish, Don't Memorize! Mitigating Memorization in Generative LLMs

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-08T10:20:02.783216Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T10:20:02.783216Z digest=sha256:d95cfb5997f9ad97021f17ae01f565386c62a687ce346e2c740d7a3f7518770e

Pith citing papers

No inbound Pith citation observations are available.