Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-06T20:36:04.931722Z
Paper Citation Record · LEDGER
As of 19 August 2026, this Paper Citation Record lists 66 of 66 outbound references and 0 inbound Pith citation observations for arXiv:2507.02456.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-06T20:36:04.931722Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-19T06:32:44.657259+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
66 of 66 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation c1d1c458-f91d-4ccf-92bb-131143fb65d8 · outbound
System-performance and cost modeling of Large Language Model training and inference Exploring the limits of transfer learning with a unified text-to-text transformer,
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9240627f-84c1-4358-a5dc-43da1e2f78de · outbound
System-performance and cost modeling of Large Language Model training and inference GLM: General language model pretraining with autoregressive blank infilling,
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 79822983-13a0-4324-8e7d-43797f51fbe6 · outbound
System-performance and cost modeling of Large Language Model training and inference BERT: Pre- training of deep bidirectional transformers for language understanding,
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 4d9fc970-16e2-4e76-8b66-4a15cee08583 · outbound
System-performance and cost modeling of Large Language Model training and inference An Image is Worth 16x16 Words: Transformers for Image Recognition at Scale
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 82037dfc-9536-426f-9681-0fb1087a8bf1 · outbound
System-performance and cost modeling of Large Language Model training and inference Training data-efficient image transformers & amp; distillation through attention,
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 2f8c24de-a924-47d2-b3b3-f2d936a81128 · outbound
System-performance and cost modeling of Large Language Model training and inference Multiple physics pretraining for physical surrogate models,
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 7ea5edc2-ac6c-457e-a784-b27440222460 · outbound
System-performance and cost modeling of Large Language Model training and inference Highly accurate protein structure prediction with alphafold,
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 6787ed9c-0e3b-49bc-b0d4-0b7c0f92a754 · outbound
System-performance and cost modeling of Large Language Model training and inference Evolutionary-scale prediction of atomic-level protein structure with a language model,
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 2ad42462-e13b-4523-b844-f38cc6e5e2c1 · outbound
System-performance and cost modeling of Large Language Model training and inference A foundation model for atomistic materials chemistry,
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation ac099849-0968-4f0c-ae4f-04218e96b6ec · outbound
System-performance and cost modeling of Large Language Model training and inference Scaling Laws for Neural Language Models
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 75b3bbd4-f5ea-4ed1-97ba-c6b8bb03c6d7 · outbound
System-performance and cost modeling of Large Language Model training and inference Efficient large-scale language model training on GPU clusters using megatron-lm,
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation beaf0e9e-52c4-4f79-94b7-be291631421a · outbound
System-performance and cost modeling of Large Language Model training and inference DeepSpeed-MoE: Advancing mixture-of- experts inference and training to power next-generation AI scale,
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 1c53770a-9cf8-41eb-a684-7a5bd99e37b1 · outbound
System-performance and cost modeling of Large Language Model training and inference DeepSpeed- inference: enabling efficient inference of transformer models at unprece- dented scale,
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 52b46adc-c68a-45d0-af41-c95df3e7785b · outbound
System-performance and cost modeling of Large Language Model training and inference OpenAI’s massive GPT-3 model is impressive, but size isn’t everything,
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 5354d26d-14b5-4dce-908a-437baeef0990 · outbound
System-performance and cost modeling of Large Language Model training and inference Flashattention: Fast and memory-efficient exact attention with io-awareness,
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 107b19e2-f183-4a68-82e2-f852051f3e64 · outbound
System-performance and cost modeling of Large Language Model training and inference FlashAttention-2: Faster Attention with Better Parallelism and Work Partitioning
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fe5b0906-b473-491a-8b3e-5a39aac6a104 · outbound
System-performance and cost modeling of Large Language Model training and inference Glam: Efficient scaling of language models with mixture-of-experts,
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 0d34d1ff-41e9-479a-94c8-75fe139c49a6 · outbound
System-performance and cost modeling of Large Language Model training and inference Hymba: A Hybrid-head Architecture for Small Language Models
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 01f52c28-0579-4c73-a4f7-e597e5724cf2 · outbound
System-performance and cost modeling of Large Language Model training and inference Flashattention-3: Fast and accurate attention with asynchrony and low- precision,
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation fec09ee5-cf3d-4057-9db9-35204a79ab1f · outbound
System-performance and cost modeling of Large Language Model training and inference From words to watts: Benchmarking the energy costs of large language model inference,
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation ac5c939c-b769-4153-841e-3a74991501c3 · outbound
System-performance and cost modeling of Large Language Model training and inference Energy- efficiency limits on training AI systems using learning-in-memory,
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 862da844-d9f3-44d2-b760-863757afa341 · outbound
System-performance and cost modeling of Large Language Model training and inference GShard: Scaling Giant Models with Conditional Computation and Automatic Sharding,
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation ddb77837-6aec-4b3c-9655-07e1d11e4b4b · outbound
System-performance and cost modeling of Large Language Model training and inference Astra-sim: Enabling sw/hw co-design exploration for distributed dl training plat- forms,
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 869ff1bc-0b50-45c8-af95-b8eb7942602a · outbound
System-performance and cost modeling of Large Language Model training and inference AI and Memory Wall ,
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 7f45dedd-f5cc-444a-89e6-0b1f4d8f1934 · outbound
System-performance and cost modeling of Large Language Model training and inference (2024, Dec) Nvidia blackwell platform arrives to power a new era of computing
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 39ba6c60-0dea-46d3-ba51-63092de6cde2 · outbound
System-performance and cost modeling of Large Language Model training and inference (2023, Jun) Amd mi300: Taming the hype - ai perfor- mance
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation c3a9a0ce-1fcc-4542-8508-badb27c8fd1c · outbound
System-performance and cost modeling of Large Language Model training and inference (2023) Why chiplets are so critical in automotive
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 7af9eb67-6a9f-4283-a5ec-bb9cdca6231b · outbound
System-performance and cost modeling of Large Language Model training and inference Performance Modeling and Workload Analysis of Distributed Large Language Model Training and Inference ,
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation ef0407cc-075e-47a2-a477-48dfa458830a · outbound
System-performance and cost modeling of Large Language Model training and inference Chiplets: How small is too small?
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 01cf57df-5953-42ba-aa88-e7e52ee402fe · outbound
System-performance and cost modeling of Large Language Model training and inference Analyzing CUDA workloads using a detailed GPU simulator,
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 08d581c8-ccdc-418b-a5fa-8c3679cb5a01 · outbound
System-performance and cost modeling of Large Language Model training and inference Accel-Sim: An extensible simulation framework for validated GPU modeling,
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8dc82372-5c50-4095-9714-53e713ed159b · outbound
System-performance and cost modeling of Large Language Model training and inference Cross- architecture performance prediction (XAPP) using CPU code to predict GPU performance,
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 14f88a40-1ea5-4dfb-a0d5-eba97493a9ec · outbound
System-performance and cost modeling of Large Language Model training and inference Principal kernel analysis: A tractable methodology to simulate scaled GPU workloads,
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 7cd64e67-eeb1-49af-bed1-a427075f48b8 · outbound
System-performance and cost modeling of Large Language Model training and inference Activation in network for NoC-based deep neural network accelerator,
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation b086559f-7855-47c8-98fa-cee930af3222 · outbound
System-performance and cost modeling of Large Language Model training and inference Performance modeling and scalability optimization of distributed deep learning systems,
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation fd91624f-6a57-4635-af9e-9b127072571c · outbound
System-performance and cost modeling of Large Language Model training and inference Paleo: A performance model for deep neural networks,
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation cbec8c55-40ee-4abb-8734-262cbafe021c · outbound
System-performance and cost modeling of Large Language Model training and inference Performance prediction of GPU- based deep learning applications,
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation c51cf814-1e0a-43f3-ba48-af894539fe4f · outbound
System-performance and cost modeling of Large Language Model training and inference Habitat: A runtime- based computational performance predictor for deep neural network training,
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 9f01cb95-8b14-498c-abaf-986755ad1b9f · outbound
System-performance and cost modeling of Large Language Model training and inference AMPeD: An analytical model for performance in dis- tributed training of transformers,
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation b31bc0f0-4566-4c55-ae70-6f2647d54ae9 · outbound
System-performance and cost modeling of Large Language Model training and inference Calculon: a methodology and tool for high-level co-design of systems and large language models,
Reference 40
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 7801bd08-b3cf-4818-a0a0-14bd3b84097c · outbound
System-performance and cost modeling of Large Language Model training and inference DeepFlow: A cross-stack pathfinding framework for distributed AI systems,
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 7a84fd28-e1a8-442b-aa66-957282cc91f9 · outbound
System-performance and cost modeling of Large Language Model training and inference Comprehensive performance modeling and system design insights for foundation models,
Reference 42
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation ad24e7ab-7403-4570-ac23-26e653e0eb0e · outbound
System-performance and cost modeling of Large Language Model training and inference A simple model for portable and fast prediction of execution time and power consumption of gpu kernels,
Reference 43
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 39cd6caf-d9a7-498e-a938-26ecadb2cd6d · outbound
System-performance and cost modeling of Large Language Model training and inference GPGPU performance and power estimation using machine learning,
Reference 44
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation f65a500d-db18-4c53-8e3e-94189737f378 · outbound
System-performance and cost modeling of Large Language Model training and inference GPU static modeling using PTX and deep structured learning,
Reference 45
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 4d62df9d-9a70-4796-a433-3a08b5934dbd · outbound
System-performance and cost modeling of Large Language Model training and inference Program analysis and machine learning–based approach to predict power consumption of cuda kernel,
Reference 46
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 5a17e420-aa07-4135-b9bd-8cfebca81794 · outbound
System-performance and cost modeling of Large Language Model training and inference Forecasting gpu performance for deep learning training and inference,
Reference 47
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation f9050f38-5820-4ca1-8617-d99107cedce4 · outbound
System-performance and cost modeling of Large Language Model training and inference Cost analysis and cost-driven IP reuse methodology for SoC design based on 2.5D/3D integration,
Reference 48
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation ad6d7033-19eb-4e31-91d0-a0111d2832f5 · outbound
System-performance and cost modeling of Large Language Model training and inference Cost-effective design of scalable high-performance systems using active and passive interposers,
Reference 49
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation dd258250-5a16-49c0-8817-a33422e867a8 · outbound
System-performance and cost modeling of Large Language Model training and inference Chiplet actuary: a quantitative cost model and multi- chiplet architecture exploration,
Reference 50
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation ed8e130f-cd72-4d62-b1cc-d1f3a73a57a6 · outbound
System-performance and cost modeling of Large Language Model training and inference Online normalizer calculation for softmax
Reference 51
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 74adcbb4-341e-4a99-a6f8-e86cb277eae9 · outbound
System-performance and cost modeling of Large Language Model training and inference From online softmax to flashattention,
Reference 52
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation ac754c03-861b-4d9e-93dd-6faf50682465 · outbound
System-performance and cost modeling of Large Language Model training and inference A hybrid tensor-expert-data parallelism approach to optimize mixture- of-experts training,
Reference 53
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation de7110b9-6ba8-4c35-9a83-a5243c12a3b5 · outbound
System-performance and cost modeling of Large Language Model training and inference Training deep learning models at scale: How nccl enables best performance on ai data center networks,
Reference 54
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 76a485d8-fc27-41f9-8129-e219edcfc1bb · outbound
System-performance and cost modeling of Large Language Model training and inference Optimization of collective communication operations in mpich,
Reference 55
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation a5e92057-f09f-4463-9089-bc5fddb9f37b · outbound
System-performance and cost modeling of Large Language Model training and inference Highly available data parallel ml training on mesh networks,
Reference 56
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation d8faf23c-3fc8-4ee4-baeb-d02d21c1e657 · outbound
System-performance and cost modeling of Large Language Model training and inference An overview of manufacturing yield and reliability modeling for semiconductor products,
Reference 57
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 7b47a4e3-3986-48ba-b069-2bc812a35f98 · outbound
System-performance and cost modeling of Large Language Model training and inference Cost-performance co- optimization for the chiplet era,
Reference 58
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 8fa55cc9-332d-4950-b6f0-110af4d847e3 · outbound
System-performance and cost modeling of Large Language Model training and inference Smoothing disruption across the stack: Tales of memory, heterogeneity, and compilers,
Reference 59
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 5d236e00-f7e8-4e74-b230-d6501449c9dd · outbound
System-performance and cost modeling of Large Language Model training and inference Eco-chip: Estimation of carbon footprint of chiplet-based architectures for sustainable vlsi,
Reference 60
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 793753d1-d306-4441-a695-ab7d05de76a1 · outbound
System-performance and cost modeling of Large Language Model training and inference Exploiting chiplet integration technol- ogy for fast high-capacity dram modules,
Reference 61
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 67af2295-df29-43e7-bf68-280fcd1e17e8 · outbound
System-performance and cost modeling of Large Language Model training and inference REED: Chiplet-Based Accelerator for Fully Homomorphic Encryption
Reference 62
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 23278291-6bae-4e82-a939-aca912426820 · outbound
System-performance and cost modeling of Large Language Model training and inference Astra-sim2.0: Modeling hierarchical networks and disaggregated systems for large-model training at scale,
Reference 63
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 5de680fd-0d9c-42a4-ba2b-7bc509b2ba18 · outbound
System-performance and cost modeling of Large Language Model training and inference (2020) Nvidia dgx a100 system architecture
Reference 64
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 6e6a679e-2a17-4eaa-93c6-0365f33b806b · outbound
System-performance and cost modeling of Large Language Model training and inference Unresolved cited work
Reference 65
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation dd93d604-a521-40d0-9f9a-9dd1cf51f009 · outbound
System-performance and cost modeling of Large Language Model training and inference A foundation model for atomistic materials chemistry
Reference 2024
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
No inbound Pith citation observations are available.