Pith. sign in

Paper Citation Record · LEDGER

System-performance and cost modeling of Large Language Model training and inference

As of 19 August 2026, this Paper Citation Record lists 66 of 66 outbound references and 0 inbound Pith citation observations for arXiv:2507.02456.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2507.02456 v1

Coverage vector

measured 66 of 66 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T20:36:04.931722Z

measured 66 of 66 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-19T06:32:44.657259+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

66 of 66 outbound references displayed

  • verified exact1
  • verified fuzzy56
  • unresolved9
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation c1d1c458-f91d-4ccf-92bb-131143fb65d8 · outbound

This paper cites Exploring the limits of transfer learning with a unified text-to-text transformer,.

System-performance and cost modeling of Large Language Model training and inference Exploring the limits of transfer learning with a unified text-to-text transformer,

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-06T20:35:59.309672Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:35:59.309672Z digest=sha256:4d8e2cda231d8f3560dddefa907db7e996b3ff44324432f1a23dc261058178ce

Observation 9240627f-84c1-4358-a5dc-43da1e2f78de · outbound

This paper cites GLM: General language model pretraining with autoregressive blank infilling,.

System-performance and cost modeling of Large Language Model training and inference GLM: General language model pretraining with autoregressive blank infilling,

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:36:15.884144Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-06T20:35:59.407398Z digest=sha256:6be5bef7b0bcb9cf084cd984a06398923c8f8918cc093823aa1fe6a6d5a14415

Observation 79822983-13a0-4324-8e7d-43797f51fbe6 · outbound

This paper cites BERT: Pre- training of deep bidirectional transformers for language understanding,.

System-performance and cost modeling of Large Language Model training and inference BERT: Pre- training of deep bidirectional transformers for language understanding,

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:36:15.676991Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-06T20:35:59.492521Z digest=sha256:b4ec11f8846d42132c9deebb42a769d119b836d997b5e37d96a0b5c2bc83493e

Observation 4d9fc970-16e2-4e76-8b66-4a15cee08583 · outbound

This paper cites An Image is Worth 16x16 Words: Transformers for Image Recognition at Scale.

System-performance and cost modeling of Large Language Model training and inference An Image is Worth 16x16 Words: Transformers for Image Recognition at Scale

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-06T20:35:59.565991Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:35:59.565991Z digest=sha256:bb24e24ea1a11d4e071b0862ede39599b74525de11a639977851e8df74a16f47

Observation 82037dfc-9536-426f-9681-0fb1087a8bf1 · outbound

This paper cites Training data-efficient image transformers & amp; distillation through attention,.

System-performance and cost modeling of Large Language Model training and inference Training data-efficient image transformers & amp; distillation through attention,

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:36:15.475069Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-06T20:35:59.661027Z digest=sha256:6d62474af03f7936c0c0de47217a42ccf5652d02ac722dbd1558321b903e15c0

Observation 2f8c24de-a924-47d2-b3b3-f2d936a81128 · outbound

This paper cites Multiple physics pretraining for physical surrogate models,.

System-performance and cost modeling of Large Language Model training and inference Multiple physics pretraining for physical surrogate models,

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:36:15.279957Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-06T20:35:59.776599Z digest=sha256:1e4ba5d07d9802da53561bb780a250ae0d0dc947dd87ceede543961ceabe6f79

Observation 7ea5edc2-ac6c-457e-a784-b27440222460 · outbound

This paper cites Highly accurate protein structure prediction with alphafold,.

System-performance and cost modeling of Large Language Model training and inference Highly accurate protein structure prediction with alphafold,

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:36:15.095937Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-06T20:35:59.882025Z digest=sha256:7019a3a58111e20ef50184385d7e92d829bd59ef98b485731f62d93664aad8fb

Observation 6787ed9c-0e3b-49bc-b0d4-0b7c0f92a754 · outbound

This paper cites Evolutionary-scale prediction of atomic-level protein structure with a language model,.

System-performance and cost modeling of Large Language Model training and inference Evolutionary-scale prediction of atomic-level protein structure with a language model,

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:36:14.907555Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-06T20:35:59.979597Z digest=sha256:8ffa7b18d21dafc7a8d1191ee430bcecd3d3386bc46fbac647a7e0074bb5dfc2

Observation 2ad42462-e13b-4523-b844-f38cc6e5e2c1 · outbound

This paper cites A foundation model for atomistic materials chemistry,.

System-performance and cost modeling of Large Language Model training and inference A foundation model for atomistic materials chemistry,

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:36:14.738188Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-06T20:36:00.100244Z digest=sha256:aeb2353381170d71ff85bbfa0c967dc7e27b092da4114cd8dab64d2351c3bcbe

Observation ac099849-0968-4f0c-ae4f-04218e96b6ec · outbound

This paper cites Scaling Laws for Neural Language Models.

System-performance and cost modeling of Large Language Model training and inference Scaling Laws for Neural Language Models

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-06T20:36:00.282601Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:36:00.282601Z digest=sha256:e39aa56732678d06551f2ee96adbde956e13e26a5f0497b85c6537d971fa70a0

Observation 75b3bbd4-f5ea-4ed1-97ba-c6b8bb03c6d7 · outbound

This paper cites Efficient large-scale language model training on GPU clusters using megatron-lm,.

System-performance and cost modeling of Large Language Model training and inference Efficient large-scale language model training on GPU clusters using megatron-lm,

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:36:14.588800Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-06T20:36:00.370532Z digest=sha256:b4b175aa90333b7dccd394c15d2764f9df17a79e78af249acc587eb4763abdda

Observation beaf0e9e-52c4-4f79-94b7-be291631421a · outbound

This paper cites DeepSpeed-MoE: Advancing mixture-of- experts inference and training to power next-generation AI scale,.

System-performance and cost modeling of Large Language Model training and inference DeepSpeed-MoE: Advancing mixture-of- experts inference and training to power next-generation AI scale,

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:36:14.337360Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-06T20:36:00.458869Z digest=sha256:f69d558548c4517a1a623eb9d87019db55e709cc31db2604f883404016b71718

Observation 1c53770a-9cf8-41eb-a684-7a5bd99e37b1 · outbound

This paper cites DeepSpeed- inference: enabling efficient inference of transformer models at unprece- dented scale,.

System-performance and cost modeling of Large Language Model training and inference DeepSpeed- inference: enabling efficient inference of transformer models at unprece- dented scale,

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:36:14.098512Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-06T20:36:00.538105Z digest=sha256:444b56a6177f442899090898832a0360d17e27fbfff88da84085ffbcd3c35288

Observation 52b46adc-c68a-45d0-af41-c95df3e7785b · outbound

This paper cites OpenAI’s massive GPT-3 model is impressive, but size isn’t everything,.

System-performance and cost modeling of Large Language Model training and inference OpenAI’s massive GPT-3 model is impressive, but size isn’t everything,

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:36:13.917167Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-06T20:36:00.604883Z digest=sha256:291fa45ce2043cb63bf4c38f498c79802bc68337ea2491f66b0701c6753dcb0a

Observation 5354d26d-14b5-4dce-908a-437baeef0990 · outbound

This paper cites Flashattention: Fast and memory-efficient exact attention with io-awareness,.

System-performance and cost modeling of Large Language Model training and inference Flashattention: Fast and memory-efficient exact attention with io-awareness,

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:36:13.728757Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-06T20:36:00.653257Z digest=sha256:449da3d1b15dbfe617c2d6196429210f117fe5d5d28fde4a1b23fd87351b5435

Observation 107b19e2-f183-4a68-82e2-f852051f3e64 · outbound

This paper cites FlashAttention-2: Faster Attention with Better Parallelism and Work Partitioning.

System-performance and cost modeling of Large Language Model training and inference FlashAttention-2: Faster Attention with Better Parallelism and Work Partitioning

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-06T20:36:00.725603Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:36:00.725603Z digest=sha256:4bb0d0436f718166a0e264add55750fa48317241142ab027231815ce4a880504

Observation fe5b0906-b473-491a-8b3e-5a39aac6a104 · outbound

This paper cites Glam: Efficient scaling of language models with mixture-of-experts,.

System-performance and cost modeling of Large Language Model training and inference Glam: Efficient scaling of language models with mixture-of-experts,

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:36:13.501073Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-06T20:36:00.784240Z digest=sha256:e75ca6ab73f8354eeb69822d35edf9c38148b01c8dd226db4832c313e54a8b84

Observation 0d34d1ff-41e9-479a-94c8-75fe139c49a6 · outbound

This paper cites Hymba: A Hybrid-head Architecture for Small Language Models.

System-performance and cost modeling of Large Language Model training and inference Hymba: A Hybrid-head Architecture for Small Language Models

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-06T20:36:00.876438Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:36:00.876438Z digest=sha256:bf6d475e684596399dadac00e560f42592aade53a0e1bb8729c77bbbe5817c55

Observation 01f52c28-0579-4c73-a4f7-e597e5724cf2 · outbound

This paper cites Flashattention-3: Fast and accurate attention with asynchrony and low- precision,.

System-performance and cost modeling of Large Language Model training and inference Flashattention-3: Fast and accurate attention with asynchrony and low- precision,

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:36:13.348016Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-06T20:36:00.991151Z digest=sha256:465caecbbe63e3ffd108f855f84413d005faa3507eeddc823785b14a47c3c19f

Observation fec09ee5-cf3d-4057-9db9-35204a79ab1f · outbound

This paper cites From words to watts: Benchmarking the energy costs of large language model inference,.

System-performance and cost modeling of Large Language Model training and inference From words to watts: Benchmarking the energy costs of large language model inference,

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:36:13.153456Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-06T20:36:01.057307Z digest=sha256:5aefa62ad43d896b0ce63359e2e8120891a7ab39674864238e5cd8e943b1d243

Observation ac5c939c-b769-4153-841e-3a74991501c3 · outbound

This paper cites Energy- efficiency limits on training AI systems using learning-in-memory,.

System-performance and cost modeling of Large Language Model training and inference Energy- efficiency limits on training AI systems using learning-in-memory,

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:36:12.996485Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-06T20:36:01.121080Z digest=sha256:9615888bb3c8ecf825e3d14546562afc4a3dfedc8c50cfe4084eb7fecfd129c9

Observation 862da844-d9f3-44d2-b760-863757afa341 · outbound

This paper cites GShard: Scaling Giant Models with Conditional Computation and Automatic Sharding,.

System-performance and cost modeling of Large Language Model training and inference GShard: Scaling Giant Models with Conditional Computation and Automatic Sharding,

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:36:12.783882Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-06T20:36:01.197895Z digest=sha256:33c1b97dc6cf80844a4646674119dffb0bc22fbf37762037b4d8966bc302f18e

Observation ddb77837-6aec-4b3c-9655-07e1d11e4b4b · outbound

This paper cites Astra-sim: Enabling sw/hw co-design exploration for distributed dl training plat- forms,.

System-performance and cost modeling of Large Language Model training and inference Astra-sim: Enabling sw/hw co-design exploration for distributed dl training plat- forms,

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:36:12.587193Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-06T20:36:01.274194Z digest=sha256:d6dd70d721b637687c0e3f868b34788eb94d479cb25e97f17ea2633d74463144

Observation 869ff1bc-0b50-45c8-af95-b8eb7942602a · outbound

This paper cites AI and Memory Wall ,.

System-performance and cost modeling of Large Language Model training and inference AI and Memory Wall ,

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:36:12.421265Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-06T20:36:01.359259Z digest=sha256:d7708553b013602e94fa8a52348e6dd4bccaede3d1a246312fc0e53758a99da0

Observation 7f45dedd-f5cc-444a-89e6-0b1f4d8f1934 · outbound

This paper cites (2024, Dec) Nvidia blackwell platform arrives to power a new era of computing.

System-performance and cost modeling of Large Language Model training and inference (2024, Dec) Nvidia blackwell platform arrives to power a new era of computing

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:36:12.218746Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-06T20:36:01.434749Z digest=sha256:0184bbceaba9dc63d0bd8156e5e283aa00dbf68e720b6ecff5067b3d3e8ddde3

Observation 39ba6c60-0dea-46d3-ba51-63092de6cde2 · outbound

This paper cites (2023, Jun) Amd mi300: Taming the hype - ai perfor- mance.

System-performance and cost modeling of Large Language Model training and inference (2023, Jun) Amd mi300: Taming the hype - ai perfor- mance

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:36:12.089157Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-06T20:36:01.513128Z digest=sha256:f9d84c267cd396232b6099339346a6fc986fec4952bc0ef6af8f337827c42cd8

Observation c3a9a0ce-1fcc-4542-8508-badb27c8fd1c · outbound

This paper cites (2023) Why chiplets are so critical in automotive.

System-performance and cost modeling of Large Language Model training and inference (2023) Why chiplets are so critical in automotive

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:36:11.948593Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-06T20:36:01.582686Z digest=sha256:244ef31cc3ef39b3eed6121ac079963c67452d4ca4cf0bcf77322ee46d0295ce

Observation 7af9eb67-6a9f-4283-a5ec-bb9cdca6231b · outbound

This paper cites Performance Modeling and Workload Analysis of Distributed Large Language Model Training and Inference ,.

System-performance and cost modeling of Large Language Model training and inference Performance Modeling and Workload Analysis of Distributed Large Language Model Training and Inference ,

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:36:11.752095Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-06T20:36:01.663839Z digest=sha256:c35a132718c4ae39fd0fe49a769fffcf860248a4f2e3388214a15dbc7638276e

Observation ef0407cc-075e-47a2-a477-48dfa458830a · outbound

This paper cites Chiplets: How small is too small?.

System-performance and cost modeling of Large Language Model training and inference Chiplets: How small is too small?

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:36:11.485003Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-06T20:36:01.722837Z digest=sha256:abc9483a3c3b2a36913ad59dec5b34ca81f649d91ad5b6e3386d0c01cb6bd60e

Observation 01cf57df-5953-42ba-aa88-e7e52ee402fe · outbound

This paper cites Analyzing CUDA workloads using a detailed GPU simulator,.

System-performance and cost modeling of Large Language Model training and inference Analyzing CUDA workloads using a detailed GPU simulator,

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:36:11.275286Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-06T20:36:01.784352Z digest=sha256:1becbc306595b75e0c31df7c25dfef4fa31770db7829435c1cc7e1a73e8a01cf

Observation 08d581c8-ccdc-418b-a5fa-8c3679cb5a01 · outbound

This paper cites Accel-Sim: An extensible simulation framework for validated GPU modeling,.

System-performance and cost modeling of Large Language Model training and inference Accel-Sim: An extensible simulation framework for validated GPU modeling,

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-06T20:36:01.855007Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:36:01.855007Z digest=sha256:e268d978caeca06cb8b07b400f00b9d280dc12feae2450dd0548e6a6487f54a1

Observation 8dc82372-5c50-4095-9714-53e713ed159b · outbound

This paper cites Cross- architecture performance prediction (XAPP) using CPU code to predict GPU performance,.

System-performance and cost modeling of Large Language Model training and inference Cross- architecture performance prediction (XAPP) using CPU code to predict GPU performance,

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:36:11.060825Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-06T20:36:01.907155Z digest=sha256:2907794709dd39bf3ce711f073cccba1f89b2a87cf2db8b0b0ccb28a6c28044f

Observation 14f88a40-1ea5-4dfb-a0d5-eba97493a9ec · outbound

This paper cites Principal kernel analysis: A tractable methodology to simulate scaled GPU workloads,.

System-performance and cost modeling of Large Language Model training and inference Principal kernel analysis: A tractable methodology to simulate scaled GPU workloads,

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:36:10.720745Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-06T20:36:02.007344Z digest=sha256:4165113ff0abc62d2d59234eaefe5aafc5416e5d9cda25cd42a944f065247e0e

Observation 7cd64e67-eeb1-49af-bed1-a427075f48b8 · outbound

This paper cites Activation in network for NoC-based deep neural network accelerator,.

System-performance and cost modeling of Large Language Model training and inference Activation in network for NoC-based deep neural network accelerator,

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:36:10.343262Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-06T20:36:02.075870Z digest=sha256:c9b11e4b2fb56f7e5a46067522c685eb9a3dea9b2b2092efe6c5bacdba9ff381

Observation b086559f-7855-47c8-98fa-cee930af3222 · outbound

This paper cites Performance modeling and scalability optimization of distributed deep learning systems,.

System-performance and cost modeling of Large Language Model training and inference Performance modeling and scalability optimization of distributed deep learning systems,

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:36:09.948280Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-06T20:36:02.185608Z digest=sha256:f0cce6cfffd4d5f052e479466333a506bd5a18484638a1b3be8c9379ae4cb6b2

Observation fd91624f-6a57-4635-af9e-9b127072571c · outbound

This paper cites Paleo: A performance model for deep neural networks,.

System-performance and cost modeling of Large Language Model training and inference Paleo: A performance model for deep neural networks,

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:36:09.574789Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-06T20:36:02.260949Z digest=sha256:a7fbb7c1f711895c9ae54172f9cd7610c62ea2a848e4d56028ff37e2848736de

Observation cbec8c55-40ee-4abb-8734-262cbafe021c · outbound

This paper cites Performance prediction of GPU- based deep learning applications,.

System-performance and cost modeling of Large Language Model training and inference Performance prediction of GPU- based deep learning applications,

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:36:09.277020Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-06T20:36:02.357371Z digest=sha256:8cfc9aaf386a9aae1b0b6e4c1118b5f9f0e57723bd167517109785783e148037

Observation c51cf814-1e0a-43f3-ba48-af894539fe4f · outbound

This paper cites Habitat: A runtime- based computational performance predictor for deep neural network training,.

System-performance and cost modeling of Large Language Model training and inference Habitat: A runtime- based computational performance predictor for deep neural network training,

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:36:09.152049Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-06T20:36:02.437306Z digest=sha256:7b599d8ba40a33fe831e77736b073f6f7ff4a5a4a99dd7ad925622cd7e9033cb

Observation 9f01cb95-8b14-498c-abaf-986755ad1b9f · outbound

This paper cites AMPeD: An analytical model for performance in dis- tributed training of transformers,.

System-performance and cost modeling of Large Language Model training and inference AMPeD: An analytical model for performance in dis- tributed training of transformers,

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:36:09.005067Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-06T20:36:02.530564Z digest=sha256:964373f8bb5c9aa7415e43ac9a839b317867f0ff73baada950e0765841629cec

Observation b31bc0f0-4566-4c55-ae70-6f2647d54ae9 · outbound

This paper cites Calculon: a methodology and tool for high-level co-design of systems and large language models,.

System-performance and cost modeling of Large Language Model training and inference Calculon: a methodology and tool for high-level co-design of systems and large language models,

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:36:08.868632Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-06T20:36:02.635168Z digest=sha256:2254d0cbec07c1691207106739e2ee820567ddf73354826b2c94066be286c5a6

Observation 7801bd08-b3cf-4818-a0a0-14bd3b84097c · outbound

This paper cites DeepFlow: A cross-stack pathfinding framework for distributed AI systems,.

System-performance and cost modeling of Large Language Model training and inference DeepFlow: A cross-stack pathfinding framework for distributed AI systems,

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:36:08.753230Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-06T20:36:02.724546Z digest=sha256:b87ddfbb5dbdbb6ff260bef19dbdcac60a43369dc85e7a87579ddbd562cb05c8

Observation 7a84fd28-e1a8-442b-aa66-957282cc91f9 · outbound

This paper cites Comprehensive performance modeling and system design insights for foundation models,.

System-performance and cost modeling of Large Language Model training and inference Comprehensive performance modeling and system design insights for foundation models,

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:36:08.611265Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-06T20:36:02.831557Z digest=sha256:4a4230fdc9079a589b1d52aaddd618b9d31a97894429a33b6f3829fc833f4f10

Observation ad24e7ab-7403-4570-ac23-26e653e0eb0e · outbound

This paper cites A simple model for portable and fast prediction of execution time and power consumption of gpu kernels,.

System-performance and cost modeling of Large Language Model training and inference A simple model for portable and fast prediction of execution time and power consumption of gpu kernels,

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:36:08.467188Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-06T20:36:02.940076Z digest=sha256:4971772f9a7ac4f6b8734f722f1e513ced64803a7559d4169aebc8c00db12eea

Observation 39cd6caf-d9a7-498e-a938-26ecadb2cd6d · outbound

This paper cites GPGPU performance and power estimation using machine learning,.

System-performance and cost modeling of Large Language Model training and inference GPGPU performance and power estimation using machine learning,

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:36:08.327262Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-06T20:36:03.016988Z digest=sha256:3f1a52a411c2c75662c11072f56d46f3918a1ff0eaed2a864ff2fd76a6b02084

Observation f65a500d-db18-4c53-8e3e-94189737f378 · outbound

This paper cites GPU static modeling using PTX and deep structured learning,.

System-performance and cost modeling of Large Language Model training and inference GPU static modeling using PTX and deep structured learning,

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:36:08.185644Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-06T20:36:03.071908Z digest=sha256:d21b864f1ca0a4c62d21b2975d53eb6828a508e071fe41b29ffd28bb74a2080d

Observation 4d62df9d-9a70-4796-a433-3a08b5934dbd · outbound

This paper cites Program analysis and machine learning–based approach to predict power consumption of cuda kernel,.

System-performance and cost modeling of Large Language Model training and inference Program analysis and machine learning–based approach to predict power consumption of cuda kernel,

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:36:08.023418Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-06T20:36:03.143503Z digest=sha256:e84d66e8691335df6978470692f4a7330b87f8c0e55658b37d698c2bd83ddcbe

Observation 5a17e420-aa07-4135-b9bd-8cfebca81794 · outbound

This paper cites Forecasting gpu performance for deep learning training and inference,.

System-performance and cost modeling of Large Language Model training and inference Forecasting gpu performance for deep learning training and inference,

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:36:07.876137Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-06T20:36:03.252009Z digest=sha256:a2bc7a17b8fd9ad99718963e60ce0d4ec8bbafa60aab0e7982f5c0a1b026f5f1

Observation f9050f38-5820-4ca1-8617-d99107cedce4 · outbound

This paper cites Cost analysis and cost-driven IP reuse methodology for SoC design based on 2.5D/3D integration,.

System-performance and cost modeling of Large Language Model training and inference Cost analysis and cost-driven IP reuse methodology for SoC design based on 2.5D/3D integration,

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:36:07.754329Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-06T20:36:03.323591Z digest=sha256:fc17aae258978059c4a3357ae4887875a682f9b6ebdd1ef864c75f0bb731b69a

Observation ad6d7033-19eb-4e31-91d0-a0111d2832f5 · outbound

This paper cites Cost-effective design of scalable high-performance systems using active and passive interposers,.

System-performance and cost modeling of Large Language Model training and inference Cost-effective design of scalable high-performance systems using active and passive interposers,

Reference 49

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:36:07.628951Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-06T20:36:03.438696Z digest=sha256:4edc77471ff83b2e4397e8a286dc518f787556ae2d60586cf1fc1d6bbaf3eca6

Observation dd258250-5a16-49c0-8817-a33422e867a8 · outbound

This paper cites Chiplet actuary: a quantitative cost model and multi- chiplet architecture exploration,.

System-performance and cost modeling of Large Language Model training and inference Chiplet actuary: a quantitative cost model and multi- chiplet architecture exploration,

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:36:07.491324Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-06T20:36:03.527251Z digest=sha256:da212d7eed60241ea95b600e7c2a956acebbe6794951e3f4b7ae6cd437d67380

Observation ed8e130f-cd72-4d62-b1cc-d1f3a73a57a6 · outbound

This paper cites Online normalizer calculation for softmax.

System-performance and cost modeling of Large Language Model training and inference Online normalizer calculation for softmax

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-06T20:36:03.647922Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:36:03.647922Z digest=sha256:abebf07c81f143c88a97cce63f2a6bfca0ece3076aa0f22fb18886986b1b7af1

Observation 74adcbb4-341e-4a99-a6f8-e86cb277eae9 · outbound

This paper cites From online softmax to flashattention,.

System-performance and cost modeling of Large Language Model training and inference From online softmax to flashattention,

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:36:07.365346Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-06T20:36:03.796628Z digest=sha256:ac0a6804939f44031c02facaa8358015c930793464ebfce07c325d9b7343adac

Observation ac754c03-861b-4d9e-93dd-6faf50682465 · outbound

This paper cites A hybrid tensor-expert-data parallelism approach to optimize mixture- of-experts training,.

System-performance and cost modeling of Large Language Model training and inference A hybrid tensor-expert-data parallelism approach to optimize mixture- of-experts training,

Reference 53

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:36:07.179821Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-06T20:36:03.850832Z digest=sha256:07428e61657778aecc3c5e5b489a5c529d9e782a82e0452165e8b9ced1bceab3

Observation de7110b9-6ba8-4c35-9a83-a5243c12a3b5 · outbound

This paper cites Training deep learning models at scale: How nccl enables best performance on ai data center networks,.

System-performance and cost modeling of Large Language Model training and inference Training deep learning models at scale: How nccl enables best performance on ai data center networks,

Reference 54

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:36:06.998039Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-06T20:36:03.930148Z digest=sha256:14e2aee6ab6cd7f5d9d178c985c511872757df39db83d3478e5bd2b2c0571204

Observation 76a485d8-fc27-41f9-8129-e219edcfc1bb · outbound

This paper cites Optimization of collective communication operations in mpich,.

System-performance and cost modeling of Large Language Model training and inference Optimization of collective communication operations in mpich,

Reference 55

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:36:06.839641Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-06T20:36:04.020247Z digest=sha256:f85115cbad8d98ea7117308a6f9de3ba182e0826ab2f0bd9103999885107566e

Observation a5e92057-f09f-4463-9089-bc5fddb9f37b · outbound

This paper cites Highly available data parallel ml training on mesh networks,.

System-performance and cost modeling of Large Language Model training and inference Highly available data parallel ml training on mesh networks,

Reference 56

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:36:06.745037Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-06T20:36:04.095555Z digest=sha256:a9decd052023d57d7e0cb965034cf6007640edbd4e0df7c3ccf6d7464428d80e

Observation d8faf23c-3fc8-4ee4-baeb-d02d21c1e657 · outbound

This paper cites An overview of manufacturing yield and reliability modeling for semiconductor products,.

System-performance and cost modeling of Large Language Model training and inference An overview of manufacturing yield and reliability modeling for semiconductor products,

Reference 57

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:36:06.639476Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-06T20:36:04.173539Z digest=sha256:5a6161a57e45ce7da85c0684235517887a99062c52bfa288ac5d928fb056519f

Observation 7b47a4e3-3986-48ba-b069-2bc812a35f98 · outbound

This paper cites Cost-performance co- optimization for the chiplet era,.

System-performance and cost modeling of Large Language Model training and inference Cost-performance co- optimization for the chiplet era,

Reference 58

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:36:06.444858Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-06T20:36:04.263853Z digest=sha256:7fd8b381d165ce0871f39e6ca81671e90f3f8661a3fe5a0b204201caa2702395

Observation 8fa55cc9-332d-4950-b6f0-110af4d847e3 · outbound

This paper cites Smoothing disruption across the stack: Tales of memory, heterogeneity, and compilers,.

System-performance and cost modeling of Large Language Model training and inference Smoothing disruption across the stack: Tales of memory, heterogeneity, and compilers,

Reference 59

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:36:06.144507Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-06T20:36:04.392126Z digest=sha256:c2e212cc08572aa36deaf83693587275fc4954994253c371c208136c24bc6341

Observation 5d236e00-f7e8-4e74-b230-d6501449c9dd · outbound

This paper cites Eco-chip: Estimation of carbon footprint of chiplet-based architectures for sustainable vlsi,.

System-performance and cost modeling of Large Language Model training and inference Eco-chip: Estimation of carbon footprint of chiplet-based architectures for sustainable vlsi,

Reference 60

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:36:05.961902Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-06T20:36:04.504890Z digest=sha256:97167cd208b5b9898ce865e8e8ba27df8563dd190f11eadf552c0c2460311fdf

Observation 793753d1-d306-4441-a695-ab7d05de76a1 · outbound

This paper cites Exploiting chiplet integration technol- ogy for fast high-capacity dram modules,.

System-performance and cost modeling of Large Language Model training and inference Exploiting chiplet integration technol- ogy for fast high-capacity dram modules,

Reference 61

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:36:05.849708Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-06T20:36:04.601202Z digest=sha256:120b76c63376a93b933c77e2b38dd61b9e494af13eda0bdf8b352da1b0570fad

Observation 67af2295-df29-43e7-bf68-280fcd1e17e8 · outbound

This paper cites REED: Chiplet-Based Accelerator for Fully Homomorphic Encryption.

System-performance and cost modeling of Large Language Model training and inference REED: Chiplet-Based Accelerator for Fully Homomorphic Encryption

Reference 62

Resolution
verified exact
local_arxiv, observed 2026-08-06T20:36:05.159919Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-06T20:36:04.657262Z digest=sha256:2ffdf8c1ccf93a6f09435750c94a365a6816568e3dadcf979d634cb4838bdc64

Observation 23278291-6bae-4e82-a939-aca912426820 · outbound

This paper cites Astra-sim2.0: Modeling hierarchical networks and disaggregated systems for large-model training at scale,.

System-performance and cost modeling of Large Language Model training and inference Astra-sim2.0: Modeling hierarchical networks and disaggregated systems for large-model training at scale,

Reference 63

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:36:05.748001Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-06T20:36:04.755119Z digest=sha256:9a63531d50664c828da43e8b76f3981138ca02288bd8d7df244e61a6a47383d2

Observation 5de680fd-0d9c-42a4-ba2b-7bc509b2ba18 · outbound

This paper cites (2020) Nvidia dgx a100 system architecture.

System-performance and cost modeling of Large Language Model training and inference (2020) Nvidia dgx a100 system architecture

Reference 64

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:36:05.559798Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-06T20:36:04.825926Z digest=sha256:bd8167919e04b0e80bb8496ca51a2a9dc678a0ac17b1c8a8d1c1c31bcfe38848

Observation 6e6a679e-2a17-4eaa-93c6-0365f33b806b · outbound

This paper cites an unresolved cited work.

System-performance and cost modeling of Large Language Model training and inference Unresolved cited work

Reference 65

Resolution
unresolved
raw_fallback, observed 2026-08-06T20:36:05.404171Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-06T20:36:04.931722Z digest=sha256:d7e595857cdae737afb999b00506fb869f91d4c365ea74565c5197b595cf147d

Observation dd93d604-a521-40d0-9f9a-9dd1cf51f009 · outbound

This paper cites A foundation model for atomistic materials chemistry.

System-performance and cost modeling of Large Language Model training and inference A foundation model for atomistic materials chemistry

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-08-06T20:36:00.175701Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:36:00.175701Z digest=sha256:89d12f47392519afba5d3411518afbd08597bb2c579b2da07e6e6e4800c7cf8c

Pith citing papers

No inbound Pith citation observations are available.