Pith. sign in

Paper Citation Record · LEDGER

Cost-Efficient LLM Training with Lifetime-Aware Tensor Offloading via GPUDirect Storage

As of 8 August 2026, this Paper Citation Record lists 39 of 39 outbound references and 0 inbound Pith citation observations for arXiv:2506.06472.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.06472 v1

Coverage vector

measured 39 of 39 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T06:00:33.427109Z

measured 39 of 39 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

39 of 39 outbound references displayed

  • verified exact0
  • verified fuzzy22
  • unresolved17
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation f2e46f6e-6bdf-45b9-938b-f02312d76d6c · outbound

This paper cites In19th USENIX Conference on File and Storage Technologies (FAST 21), pages 387–401, 2021.

Cost-Efficient LLM Training with Lifetime-Aware Tensor Offloading via GPUDirect Storage In19th USENIX Conference on File and Storage Technologies (FAST 21), pages 387–401, 2021

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T06:00:38.224953Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T06:00:32.204819Z digest=sha256:e982f141f82406cea306ed85f705efba457745d3009e12518d348a2542da5f7b

Observation 8b9ac2a1-8192-4455-a34a-31ae374be6b4 · outbound

This paper cites Efficient combination of rematerialization and offloading for training dnns.Advances in Neural Information Processing Systems, 34:23844–23857, 2021.

Cost-Efficient LLM Training with Lifetime-Aware Tensor Offloading via GPUDirect Storage Efficient combination of rematerialization and offloading for training dnns.Advances in Neural Information Processing Systems, 34:23844–23857, 2021

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T06:00:38.076966Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T06:00:32.210913Z digest=sha256:ad226121971b38be16bc560029b068b33204fd3b4d684153aa909ff5b242a23c

Observation 75a2931f-a17e-40b1-a1bb-47b4253cdbf5 · outbound

This paper cites BERT: Pre-training of Deep Bidirectional Transformers for Language Understanding.

Cost-Efficient LLM Training with Lifetime-Aware Tensor Offloading via GPUDirect Storage BERT: Pre-training of Deep Bidirectional Transformers for Language Understanding

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-07T06:00:32.219076Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T06:00:32.219076Z digest=sha256:de5ea0da887f1b4bd25af359f70b8343f534021dab850ef71a0976664d904830

Observation 4010036c-c454-4239-ae2b-b2339ce45d89 · outbound

This paper cites The Llama 3 Herd of Models.

Cost-Efficient LLM Training with Lifetime-Aware Tensor Offloading via GPUDirect Storage The Llama 3 Herd of Models

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-07T06:00:32.223943Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T06:00:32.223943Z digest=sha256:01292c4a9811811f10f8ceade8bbc4bc112cd87c420a0f5b7644cc50205b8988

Observation 157e3f5e-29db-458b-8529-bd6344290d2c · outbound

This paper cites Exxact.https://www.exxactcorp.com/, 2025.

Cost-Efficient LLM Training with Lifetime-Aware Tensor Offloading via GPUDirect Storage Exxact.https://www.exxactcorp.com/, 2025

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T06:00:37.885091Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T06:00:32.229730Z digest=sha256:5f7161d91ebf56cabca11aefdcd5ae3a076f1dc1e436904345e84853a4388daf

Observation 79269ad2-b570-4a8c-a116-40c88095f66c · outbound

This paper cites T5 11b.https://huggingface.co/google-t5/t5-11b, 2025.

Cost-Efficient LLM Training with Lifetime-Aware Tensor Offloading via GPUDirect Storage T5 11b.https://huggingface.co/google-t5/t5-11b, 2025

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T06:00:37.668737Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T06:00:32.234022Z digest=sha256:dce513db8cee305c26e438cb1c0713c4a101e870b38da77c958bef3181d21bbe

Observation 0220b629-eb06-4ca6-af42-eb96375c0285 · outbound

This paper cites nvidia.com/blog/gpudirect-storage/.

Cost-Efficient LLM Training with Lifetime-Aware Tensor Offloading via GPUDirect Storage nvidia.com/blog/gpudirect-storage/

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T06:00:37.498608Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T06:00:32.244791Z digest=sha256:5af89811eaee525bbc3b693b85133d72ff4dfbd2577c50ef4157513907b88ac1

Observation fa229a1c-7736-48aa-901c-63497343213c · outbound

This paper cites Swapadvisor: Pushingdeeplearningbeyondthegpu memory limit via smart swapping.

Cost-Efficient LLM Training with Lifetime-Aware Tensor Offloading via GPUDirect Storage Swapadvisor: Pushingdeeplearningbeyondthegpu memory limit via smart swapping

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T06:00:37.315978Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T06:00:32.249225Z digest=sha256:60f7957ce6aba6896004757f9fbeefb28f013e4d5f989193a7a74dbd41bc7691

Observation 565f4070-ef28-41fb-a56a-b5144f07d578 · outbound

This paper cites Gpipe: Efficient training of giant neural networks using pipeline parallelism.Advances in neural information processing systems, 32, 2019.

Cost-Efficient LLM Training with Lifetime-Aware Tensor Offloading via GPUDirect Storage Gpipe: Efficient training of giant neural networks using pipeline parallelism.Advances in neural information processing systems, 32, 2019

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-07T06:00:32.255025Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T06:00:32.255025Z digest=sha256:58a4f0dfc426934ca2f5767cd71d5fe90263347fcf5626c347dc2347caba21fa

Observation 90eeb0ea-75ae-43c7-9042-18b66e18393f · outbound

This paper cites Ibm granite.https://huggingface.co/ibm-granite, 2025.

Cost-Efficient LLM Training with Lifetime-Aware Tensor Offloading via GPUDirect Storage Ibm granite.https://huggingface.co/ibm-granite, 2025

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T06:00:37.132856Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T06:00:32.262423Z digest=sha256:5360122e7d52bd27a53608cef2a4232915b9206454720ae56c8ecea5f5116dfb

Observation 05d469d2-ce6e-4e5c-8cd2-08bfb9f65c09 · outbound

This paper cites Checkmate: Breaking the memory wall with optimal tensor rematerialization.Proceedings of Machine Learning and Systems, 2:497–511, 2020.

Cost-Efficient LLM Training with Lifetime-Aware Tensor Offloading via GPUDirect Storage Checkmate: Breaking the memory wall with optimal tensor rematerialization.Proceedings of Machine Learning and Systems, 2:497–511, 2020

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T06:00:36.954625Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T06:00:32.271033Z digest=sha256:6c60b547da68e5c080974a7bdfa100abc3816c857ca752f44f75827edd62fa38

Observation 66b4a927-a40b-4dd6-84cc-2fc2347ddce8 · outbound

This paper cites Smart-infinity: Fastlargelanguagemodeltrainingusingnear-storageprocessingonarealsystem.

Cost-Efficient LLM Training with Lifetime-Aware Tensor Offloading via GPUDirect Storage Smart-infinity: Fastlargelanguagemodeltrainingusingnear-storageprocessingonarealsystem

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T06:00:36.734315Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T06:00:32.281042Z digest=sha256:9236262b2a3616d814cf0fb3bc82e400bfd5abf3f6bd23fdc7821e97e9b74e0a

Observation 62e93b52-97ee-480b-9a82-f474a889d2ba · outbound

This paper cites Deepum: Tensormigrationandprefetchinginunified memory.

Cost-Efficient LLM Training with Lifetime-Aware Tensor Offloading via GPUDirect Storage Deepum: Tensormigrationandprefetchinginunified memory

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T06:00:36.517578Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T06:00:32.291031Z digest=sha256:74c8e8c3b162deeeded6b50ab47504c82d68be63117bfcc8fe85e21c457dc7e4

Observation 9e4d01fd-f963-4c16-aac4-afdd38782399 · outbound

This paper cites Scaling Laws for Neural Language Models.

Cost-Efficient LLM Training with Lifetime-Aware Tensor Offloading via GPUDirect Storage Scaling Laws for Neural Language Models

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-07T06:00:32.301738Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T06:00:32.301738Z digest=sha256:4802068eaf2b6b1b97c7d098001b6a6fb025c052b2f853d0773d191bdec871d4

Observation 1ee83164-88a9-44c8-8c30-d5f8387bc992 · outbound

This paper cites Beyond the memory wall: A case for memory-centric hpc system for deep learning.

Cost-Efficient LLM Training with Lifetime-Aware Tensor Offloading via GPUDirect Storage Beyond the memory wall: A case for memory-centric hpc system for deep learning

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T06:00:36.317589Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T06:00:32.312924Z digest=sha256:76b622f85b0dc6f77f76d9a606cff5c5215c28fb1a489b9ae18196d2e1413f5f

Observation 25d743ac-8fb4-4732-845d-9763d4d96a9a · outbound

This paper cites TFLMS: Large Model Support in TensorFlow by Graph Rewriting.

Cost-Efficient LLM Training with Lifetime-Aware Tensor Offloading via GPUDirect Storage TFLMS: Large Model Support in TensorFlow by Graph Rewriting

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-07T06:00:32.323175Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T06:00:32.323175Z digest=sha256:797e3d6957a90fbf8fb9f7279890f7ab60ec90987828c1b475d8eaa46163ff64

Observation 6c2374d8-4fe3-4c22-b633-27febef3d874 · outbound

This paper cites TorchTitan: One-stop PyTorch native solution for production ready LLM pre-training.

Cost-Efficient LLM Training with Lifetime-Aware Tensor Offloading via GPUDirect Storage TorchTitan: One-stop PyTorch native solution for production ready LLM pre-training

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-07T06:00:32.332267Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T06:00:32.332267Z digest=sha256:63b0458a3ed88111744bd2998c30df5dd40f0d86b26ec82ca898f0c804e30378

Observation 9347a12c-dc38-4126-87ec-492546f5dec0 · outbound

This paper cites LoHan: Low-Cost High-Performance Framework to Fine-Tune 100B Model on a Consumer GPU.

Cost-Efficient LLM Training with Lifetime-Aware Tensor Offloading via GPUDirect Storage LoHan: Low-Cost High-Performance Framework to Fine-Tune 100B Model on a Consumer GPU

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-07T06:00:32.342788Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T06:00:32.342788Z digest=sha256:095eb050d0e224793c2125291aa37e00b9fbe2f81ae5305d1e21391b8bf5045e

Observation 93c3f85c-3e80-4487-aace-506d2cf81670 · outbound

This paper cites An Empirical Model of Large-Batch Training.

Cost-Efficient LLM Training with Lifetime-Aware Tensor Offloading via GPUDirect Storage An Empirical Model of Large-Batch Training

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-07T06:00:32.354308Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T06:00:32.354308Z digest=sha256:f6da1bed43d541245f9de1ae3ddcc462d37457b9e7736ce93086e02353574fa4

Observation 6b5ea083-c2da-4c5a-a012-ecf0710c1997 · outbound

This paper cites Mixed Precision Training.

Cost-Efficient LLM Training with Lifetime-Aware Tensor Offloading via GPUDirect Storage Mixed Precision Training

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-07T06:00:32.365297Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T06:00:32.365297Z digest=sha256:cd3bb713ece586f9f8b947a7cf205852330a807f26e6d748eb4ddea521cc7879

Observation c7725067-0716-481a-ba8e-9b15d1ccdbdd · outbound

This paper cites Pipedream: Generalized pipeline parallelism for dnn training.

Cost-Efficient LLM Training with Lifetime-Aware Tensor Offloading via GPUDirect Storage Pipedream: Generalized pipeline parallelism for dnn training

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-07T06:00:32.377846Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T06:00:32.377846Z digest=sha256:91132078d280f385756a342440362138586f7efbddb2a2682409bed17576dd65

Observation 4244f784-58d3-4f9b-b7ce-5b136bfbc231 · outbound

This paper cites Angel-PTM: A Scalable and Economical Large-scale Pre-training System in Tencent.

Cost-Efficient LLM Training with Lifetime-Aware Tensor Offloading via GPUDirect Storage Angel-PTM: A Scalable and Economical Large-scale Pre-training System in Tencent

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-07T06:00:32.388782Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T06:00:32.388782Z digest=sha256:326b8766953b71d99b22a8694e56589f4f7ee301bee42b8fce107edf27d5cc87

Observation e21e5ca9-0316-495e-81fd-55617bd9174e · outbound

This paper cites Automatic differentiation in pytorch.

Cost-Efficient LLM Training with Lifetime-Aware Tensor Offloading via GPUDirect Storage Automatic differentiation in pytorch

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-07T06:00:32.401389Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T06:00:32.401389Z digest=sha256:809e14fdc7ff01e036743f8566e91be48e07306051cc8505ed735a0257ed0a57

Observation 0f628802-5c38-4ba1-8595-23609fd97d9c · outbound

This paper cites Scalable diffusion models with transformers.

Cost-Efficient LLM Training with Lifetime-Aware Tensor Offloading via GPUDirect Storage Scalable diffusion models with transformers

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-07T06:00:32.412497Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T06:00:32.412497Z digest=sha256:209044bbca18a7a1b0d4b5e6989800d97a9895f2128a36aa780d7df97a27c713

Observation 14a0829f-bdb2-48d5-916a-ca38b86e1ebd · outbound

This paper cites Training Large Neural Networks with Constant Memory using a New Execution Algorithm.

Cost-Efficient LLM Training with Lifetime-Aware Tensor Offloading via GPUDirect Storage Training Large Neural Networks with Constant Memory using a New Execution Algorithm

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-07T06:00:32.422749Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T06:00:32.422749Z digest=sha256:cd09eea54bf1ff58292337a44a9734a22b07199b30a4669f3feb5a7f21272a01

Observation 5047daf3-6af8-4ed3-8baa-742822822aaf · outbound

This paper cites Gpu-initiated on-demand high-throughput storage access in the bam system architecture.

Cost-Efficient LLM Training with Lifetime-Aware Tensor Offloading via GPUDirect Storage Gpu-initiated on-demand high-throughput storage access in the bam system architecture

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T06:00:36.064077Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T06:00:32.431974Z digest=sha256:ef1bfc9e2632ccd34bcc8310255fd98ece1a382cbca467d4627d6fbd577f1228

Observation 635b8491-36e0-499a-ae3f-3c6646cb1c40 · outbound

This paper cites Language models are unsupervised multitask learners.OpenAI blog, 1(8):9, 2019.

Cost-Efficient LLM Training with Lifetime-Aware Tensor Offloading via GPUDirect Storage Language models are unsupervised multitask learners.OpenAI blog, 1(8):9, 2019

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-07T06:00:32.441563Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T06:00:32.441563Z digest=sha256:9a1e1c70759e79baa7c6943704d53b24e37dadeeefa1923f4c09d8a71a2fd694

Observation 01ad0130-5142-4528-bb29-3d84f7a75259 · outbound

This paper cites Exploring the limits of transfer learning with a unified text-to-text transformer.Journal of machine learning research, 21(140):1–67, 2020.

Cost-Efficient LLM Training with Lifetime-Aware Tensor Offloading via GPUDirect Storage Exploring the limits of transfer learning with a unified text-to-text transformer.Journal of machine learning research, 21(140):1–67, 2020

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-07T06:00:32.485821Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T06:00:32.485821Z digest=sha256:405e4b6c7f4d6dd914141c14463a0c742da8d23b656af6891865e265db6e3231

Observation ea7d54cc-26f5-4bec-8c70-07fc2c4a603c · outbound

This paper cites Zero- infinity: Breaking the gpu memory wall for extreme scale deep learning.

Cost-Efficient LLM Training with Lifetime-Aware Tensor Offloading via GPUDirect Storage Zero- infinity: Breaking the gpu memory wall for extreme scale deep learning

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T06:00:35.776160Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T06:00:32.546527Z digest=sha256:9294fc6883a8cfa203bacaf82408a2f9169ba8e7c4b57e6361ebf32f76378c3b

Observation 60312ea8-13a6-49bd-acc7-eda91f132159 · outbound

This paper cites {Zero-offload}: Democratizing{billion-scale}model training.

Cost-Efficient LLM Training with Lifetime-Aware Tensor Offloading via GPUDirect Storage {Zero-offload}: Democratizing{billion-scale}model training

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T06:00:35.580069Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T06:00:32.628054Z digest=sha256:c86c8eb787582d662e4839eebe7e7fee35797b0656a60114d1559e43167876c8

Observation 612eff07-8439-4277-b73d-1afec967f1e8 · outbound

This paper cites vdnn: Virtualizeddeepneuralnetworksforscalable,memory-efficientneuralnetworkdesign.

Cost-Efficient LLM Training with Lifetime-Aware Tensor Offloading via GPUDirect Storage vdnn: Virtualizeddeepneuralnetworksforscalable,memory-efficientneuralnetworkdesign

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T06:00:35.359528Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T06:00:32.723638Z digest=sha256:89a2b85c6364ec0fce222626790dbe818854c725011f03f867f9e591ca1709e3

Observation fe8d81db-435b-4143-9666-f0460fa68928 · outbound

This paper cites Building ai agents for autonomous clouds: Challenges and design principles.

Cost-Efficient LLM Training with Lifetime-Aware Tensor Offloading via GPUDirect Storage Building ai agents for autonomous clouds: Challenges and design principles

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T06:00:35.090008Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T06:00:32.812774Z digest=sha256:880d13916976c53627c78fb8096225bce18119b6b6589e37bf6065bf99e6fdb8

Observation 4899a902-1259-4f04-a872-4a13a2e6d2b3 · outbound

This paper cites Stronghold: fast and affordable billion-scale deep learning model training.

Cost-Efficient LLM Training with Lifetime-Aware Tensor Offloading via GPUDirect Storage Stronghold: fast and affordable billion-scale deep learning model training

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T06:00:34.872110Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T06:00:32.885983Z digest=sha256:01fd224d449fb5909f8d693c82fd9a40ad806bf07a71758cdf89d8f09beab986

Observation 87b62046-85c8-4127-ae0d-7ee6f187fbb5 · outbound

This paper cites Superneurons: Dynamic gpu memory management for training deep neural networks.

Cost-Efficient LLM Training with Lifetime-Aware Tensor Offloading via GPUDirect Storage Superneurons: Dynamic gpu memory management for training deep neural networks

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T06:00:34.654946Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T06:00:32.986663Z digest=sha256:7e4b32a2a6892916fa1043ed68d5520f79ae377b178db6e21a09c0a3b6fce03c

Observation 60397c44-59f8-443c-abe6-5f9a2c279258 · outbound

This paper cites Netllm: Adapting large language models for networking.

Cost-Efficient LLM Training with Lifetime-Aware Tensor Offloading via GPUDirect Storage Netllm: Adapting large language models for networking

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T06:00:34.439392Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T06:00:33.078030Z digest=sha256:513a120ba8cf5d3bbcfb52e870549b1cd70527195c16e7021ab9c86a68c66ac8

Observation 108aa304-7f91-4023-931c-365f0d00c4d5 · outbound

This paper cites SSDTrain: An Activation Offloading Framework to SSDs for Faster Large Language Model Training.

Cost-Efficient LLM Training with Lifetime-Aware Tensor Offloading via GPUDirect Storage SSDTrain: An Activation Offloading Framework to SSDs for Faster Large Language Model Training

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-07T06:00:33.164055Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T06:00:33.164055Z digest=sha256:a535a660fd48755480212534757c243f45d43cb9540695850199c3491324f8ca

Observation 57779a89-1c9f-4080-9835-0f54f9743c4d · outbound

This paper cites Acceleratingthetrainingoflargelanguagemodelsusingefficientactivation rematerializationandoptimalhybridparallelism.

Cost-Efficient LLM Training with Lifetime-Aware Tensor Offloading via GPUDirect Storage Acceleratingthetrainingoflargelanguagemodelsusingefficientactivation rematerializationandoptimalhybridparallelism

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T06:00:34.182335Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T06:00:33.266662Z digest=sha256:cfd797f3878ba0990a5a8ad4ce6d20a151a42d8cd5826bcd2cd195022e188b44

Observation ca18bf44-aead-4cb4-bb78-9fb324554960 · outbound

This paper cites Zng: Architecting gpu multi-processors with new flash for scalable data analysis.

Cost-Efficient LLM Training with Lifetime-Aware Tensor Offloading via GPUDirect Storage Zng: Architecting gpu multi-processors with new flash for scalable data analysis

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T06:00:34.008635Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T06:00:33.346043Z digest=sha256:978cee0e99787872786001d18351d07ba426d4ac1ee637eec0adf4a469c44f8d

Observation f5f669cb-a2c0-4bdd-b073-8e9a4d7e7030 · outbound

This paper cites Flashgpu: Placing new flash next to gpu cores.

Cost-Efficient LLM Training with Lifetime-Aware Tensor Offloading via GPUDirect Storage Flashgpu: Placing new flash next to gpu cores

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T06:00:33.797888Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T06:00:33.427109Z digest=sha256:8cc3815a79616a245777025fbb392eaec46d13e6bc0a53ce4399753f956d84e7

Pith citing papers

No inbound Pith citation observations are available.