Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-06T15:24:17.854387Z
Paper Citation Record · LEDGER
As of 7 August 2026, this Paper Citation Record lists 15 of 15 outbound references and 8 inbound Pith citation observations for arXiv:2507.16099.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-06T15:24:17.854387Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-01T15:11:09.157452Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-07-04T06:29:37.605132Z
15 of 15 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 6ac0db3a-5086-4245-be3f-038de41f42d0 · outbound
TorchAO: PyTorch-Native Training-to-Serving Model Optimization Accelerating Transformer Inference and Training with 2:4 Activation Sparsity
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 564d2034-e16c-47d9-8aa3-0f0cbb3509e4 · outbound
TorchAO: PyTorch-Native Training-to-Serving Model Optimization PARQ: Piecewise-Affine Regularized Quantization
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b832b031-4885-4357-8925-1d1b508f3784 · outbound
TorchAO: PyTorch-Native Training-to-Serving Model Optimization TorchTitan: One-stop PyTorch native solution for production ready LLM pre-training
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 28dc0925-555a-4564-98fb-2ff809e8a9e7 · outbound
TorchAO: PyTorch-Native Training-to-Serving Model Optimization SpinQuant: LLM quantization with learned rotations
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 30b9630a-19d8-4f88-b4d8-f816bd38a0c7 · outbound
TorchAO: PyTorch-Native Training-to-Serving Model Optimization Paretoq: Scaling laws in extremely low-bit llm quantization
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 74821af5-3621-4e38-a94a-85b7a787aef5 · outbound
TorchAO: PyTorch-Native Training-to-Serving Model Optimization Accelerating Sparse Deep Neural Networks
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b732f6c7-90c8-4a87-af9b-f6255f175922 · outbound
TorchAO: PyTorch-Native Training-to-Serving Model Optimization Microscaling Data Formats for Deep Learning
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c6293d38-3f58-4ad3-a320-9682b93cfbd1 · outbound
TorchAO: PyTorch-Native Training-to-Serving Model Optimization HuggingFace's Transformers: State-of-the-art Natural Language Processing
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 66dda01f-7d0e-465e-9fa7-60b22784bcb4 · outbound
TorchAO: PyTorch-Native Training-to-Serving Model Optimization GaLore: Memory-Efficient LLM Training by Gradient Low-Rank Projection
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ebc71b69-400d-4859-bc55-e642d9b44877 · outbound
TorchAO: PyTorch-Native Training-to-Serving Model Optimization SGLang: Efficient Execution of Structured Language Model Programs
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9b71d901-8d56-41e2-bc1f-11bebb7fc930 · outbound
TorchAO: PyTorch-Native Training-to-Serving Model Optimization Torchao: Low-bit arm cpu and metal ker- nels for linear and embedding ops
Reference 2021
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation f1902be7-ee35-4736-88f9-ce89c36aae13 · outbound
TorchAO: PyTorch-Native Training-to-Serving Model Optimization Llama Guard: LLM-based Input-Output Safeguard for Human-AI Conversations
Reference 2022
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c3f3c4be-15f1-46e3-a09e-4fea89dbad4a · outbound
TorchAO: PyTorch-Native Training-to-Serving Model Optimization LLM.int8(): 8-bit Matrix Multiplication for Transformers at Scale
Reference 2023
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9d3bdde2-b7e0-4445-bd84-c6aaa43f1abe · outbound
TorchAO: PyTorch-Native Training-to-Serving Model Optimization DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning
Reference 2024
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3ac0f1cc-33d8-4b9f-a5b7-b6e9d9b4f38c · outbound
TorchAO: PyTorch-Native Training-to-Serving Model Optimization The Llama 3 Herd of Models
Reference 2025
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cc0bb48c-2b7f-4807-b097-a7f8e6b65a48 · inbound
Zero-Shot Quantization via Weight-Space Arithmetic TorchAO: PyTorch-Native Training-to-Serving Model Optimization
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 7c13a4b4-3f55-48c8-a253-9f9f7be4021e · inbound
StoSignSGD: Unbiased Structural Stochasticity Fixes SignSGD for Training Large Language Models TorchAO: PyTorch-Native Training-to-Serving Model Optimization
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 159391d1-f2ed-4895-9a8d-0999c5284a78 · inbound
LoKA: Low-precision Kernel Applications for Recommendation Models At Scale TorchAO: PyTorch-Native Training-to-Serving Model Optimization
Reference 64
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation e25d6f60-c26b-4a7a-8a7e-5ec5fe1938a5 · inbound
LoKA: Low-precision Kernel Applications for Recommendation Models At Scale TorchAO: PyTorch-Native Training-to-Serving Model Optimization
Reference 64
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 931e6dab-9c33-4ac9-8d17-87802f39186a · inbound
torchtune: PyTorch native post-training library TorchAO: PyTorch-Native Training-to-Serving Model Optimization
Reference 62
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 74c80bb9-a86d-47ca-a07a-74b108be140b · inbound
CAT-Translate: Building Compact Open-Source Models for Japanese-English Translation TorchAO: PyTorch-Native Training-to-Serving Model Optimization
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 136220bb-1fef-44b7-9aba-4e218fb10ad2 · inbound
StreamDQ: Near-Memory Weight DeQuantization in Custom HBM for Scalable AI Inference Acceleration TorchAO: PyTorch-Native Training-to-Serving Model Optimization
Reference 52
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1b1ccc5c-42f3-402b-8bb4-12c1d15be4f8 · inbound
Recti-Q: Feature-Space Rectification for Out-of-Distribution-Robust Quantized Perception in Edge Robotics TorchAO: PyTorch-Native Training-to-Serving Model Optimization
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.