Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 10 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 19 inbound Pith citation observations for arXiv:2409.12517.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-09T11:23:40.929720Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-07-03T07:27:44.431731Z
0 of 0 outbound references displayed
External citation measurements
No source-named external measurement is stored.
No outbound reference observations are available for this paper version.
Observation c4b9be6e-0317-4925-8dc5-709c656c726e · inbound
NVILA: Efficient Frontier Visual Language Models Scaling FP8 training to trillion-token LLMs
Reference 86
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation f804b551-985d-46b4-9dc7-9f50336ff2f0 · inbound
Peri-LN: Revisiting Normalization Layer in the Transformer Architecture Scaling FP8 training to trillion-token LLMs
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 67aba63f-2798-4df1-9f36-1e7f316667f5 · inbound
QuantSpec: Self-Speculative Decoding with Hierarchical Quantized KV Cache Scaling FP8 training to trillion-token LLMs
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4cac8f33-ef05-41ca-aeeb-e6fc175a5a72 · inbound
Scaling Law for Quantization-Aware Training Scaling FP8 training to trillion-token LLMs
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f59b69c3-0296-407c-b2e7-12e45dd8d740 · inbound
FP4 All the Way: Fully Quantized Training of LLMs Scaling FP8 training to trillion-token LLMs
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6873f7de-f463-4129-8d2a-ad77c575243b · inbound
Recipes for Pre-training LLMs with MXFP8 Scaling FP8 training to trillion-token LLMs
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 87602fef-a9ad-4628-a28a-e57900c9ba43 · inbound
Characterization and Mitigation of Training Instabilities in Microscaling Formats Scaling FP8 training to trillion-token LLMs
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d5b58414-1fa1-4c1b-b30d-a0aa36c878c4 · inbound
Thunder-LLM: Efficiently Adapting LLMs to Korean with Minimal Resources Scaling FP8 training to trillion-token LLMs
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0ef52389-2791-4e33-9958-2e8aac0f444e · inbound
DeepSeek: Paradigm Shifts and Technical Evolution in Large AI Models Scaling FP8 training to trillion-token LLMs
Reference 128
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a17c0e55-6fb2-471e-9476-971700957f1f · inbound
A Comprehensive FP8 Training Recipe for Reasoning-Enhanced Language Models Scaling FP8 training to trillion-token LLMs
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2747e26a-663e-407e-9239-46432204b867 · inbound
Why Low-Precision Transformer Training Fails: An Analysis on Flash Attention Scaling FP8 training to trillion-token LLMs
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 16d8a70c-f9d1-4dae-b920-fb03d4795f40 · inbound
From Detection to Recovery: Operational Analysis on LLM Pre-training with 504 GPUs Scaling FP8 training to trillion-token LLMs
Reference 49
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation ceea8870-8af1-47ab-941f-aaac89e01fb9 · inbound
From Detection to Recovery: Operational Analysis on LLM Pre-training with 504 GPUs Scaling FP8 training to trillion-token LLMs
Reference 49
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 8732e7ae-e736-45f3-b579-645a99dd585c · inbound
LoKA: Low-precision Kernel Applications for Recommendation Models At Scale Scaling FP8 training to trillion-token LLMs
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation aee8c9c7-516b-4276-869f-b4e7a917cc26 · inbound
LoKA: Low-precision Kernel Applications for Recommendation Models At Scale Scaling FP8 training to trillion-token LLMs
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation ec427814-b049-427b-889c-79e1561f93b9 · inbound
Expand More, Shrink Less: Shaping Effective-Rank Dynamics for Dense Scaling in Recommendation Scaling FP8 training to trillion-token LLMs
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 04e1bcac-80e3-4b73-b29a-26833937f021 · inbound
Achieving Cloud-Grade SLOs for Local Mixture-of-Experts Inference through CPU-GPU Hybrid Design Scaling FP8 training to trillion-token LLMs
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 7ae5b301-63ae-49e1-8d4e-0c59f242f88d · inbound
Full-Stack FP4: Stable LLM Pretraining with Quantized Projections, Optimizers, and Attention Scaling FP8 training to trillion-token LLMs
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fc699b5a-1423-467d-977a-07cd651a73b9 · inbound
One QK Channel, Many Sources: Guarding Low-Precision Attention Collapse Scaling FP8 training to trillion-token LLMs
Reference 2024
Source-reported events for the cited work
Unavailable: canonical work link unavailable.