Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-09T21:31:36.314665Z
Paper Citation Record · LEDGER
As of 17 August 2026, this Paper Citation Record lists 47 of 47 outbound references and 2 inbound Pith citation observations for arXiv:2501.19090.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-09T21:31:36.314665Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-17T06:30:58.91139+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-14T04:38:43.574201Z
A source-named dated measurement, never combined with another source.
Source: pith, observed 2026-08-14T04:38:46.369213Z
47 of 47 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation e4848237-806d-4074-8f1d-8e029f5dc6e6 · outbound
Pivoting Factorization: A Compact Meta Low-Rank Representation of Sparsity for Efficient Inference in Large Language Models URL https://www.nvidia.com/content/dam/en-zz/Solutions/Data-Center/nvidia-ampere-architecture-whitepaper.pdf
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 5f53974b-42b5-4c3e-87c5-382631f4d5fc · outbound
Pivoting Factorization: A Compact Meta Low-Rank Representation of Sparsity for Efficient Inference in Large Language Models L., do Nascimento, M
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6d6ed2b4-c2dd-4942-8571-f25250179103 · outbound
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 6949f443-542d-4423-b5cf-6b106d9e6ed2 · outbound
Pivoting Factorization: A Compact Meta Low-Rank Representation of Sparsity for Efficient Inference in Large Language Models Beyond Size: How Gradients Shape Pruning Decisions in Large Language Models
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 17896099-be58-48fb-9f6e-2953119a7634 · outbound
Pivoting Factorization: A Compact Meta Low-Rank Representation of Sparsity for Efficient Inference in Large Language Models Pruner-zero: Evolving symbolic pruning metric from scratch for large language models
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation ca7173d3-b69c-4c7d-ae14-6a97490fc23a · outbound
Pivoting Factorization: A Compact Meta Low-Rank Representation of Sparsity for Efficient Inference in Large Language Models The Llama 3 Herd of Models
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 27b55f0c-88bb-4b1c-81a9-a9451fa5c58e · outbound
Pivoting Factorization: A Compact Meta Low-Rank Representation of Sparsity for Efficient Inference in Large Language Models Mask LLM : Learnable semi-structured sparsity for large language models
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation b3a08527-0fe7-4e02-8c5f-2e85f6c9c172 · outbound
Pivoting Factorization: A Compact Meta Low-Rank Representation of Sparsity for Efficient Inference in Large Language Models and Alistarh, D
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b3faf3f0-007f-4545-9e85-6002ed56a5fc · outbound
Pivoting Factorization: A Compact Meta Low-Rank Representation of Sparsity for Efficient Inference in Large Language Models A framework for few-shot language model evaluation, 12 2023
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cd108b60-83a6-4947-87be-bffbd7ec3dac · outbound
Pivoting Factorization: A Compact Meta Low-Rank Representation of Sparsity for Efficient Inference in Large Language Models Disp-llm: Dimension-independent structural pruning for large language models
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 5b38f141-f285-418f-a966-6f074102352c · outbound
Pivoting Factorization: A Compact Meta Low-Rank Representation of Sparsity for Efficient Inference in Large Language Models G., and Wolff, G
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation fe9858e1-adec-4211-8e16-88bb2f48554b · outbound
Pivoting Factorization: A Compact Meta Low-Rank Representation of Sparsity for Efficient Inference in Large Language Models Language model compression with weighted low-rank factorization
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4cbe86d9-0b29-495b-aac8-e374e3dd741b · outbound
Pivoting Factorization: A Compact Meta Low-Rank Representation of Sparsity for Efficient Inference in Large Language Models From Low Rank Gradient Subspace Stabilization to Low-Rank Weights: Observations, Theories, and Applications
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3713a045-c1e5-4660-8a67-f72d086b1cc7 · outbound
Pivoting Factorization: A Compact Meta Low-Rank Representation of Sparsity for Efficient Inference in Large Language Models LORD: Low Rank Decomposition Of Monolingual Code LLMs For One-Shot Compression
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0f16a9fa-c651-46b8-8123-13757b423103 · outbound
Pivoting Factorization: A Compact Meta Low-Rank Representation of Sparsity for Efficient Inference in Large Language Models Optimal brain damage, advances in neural information processing systems
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation fbd2ba6a-479f-4095-b7a7-cf042ce5eb75 · outbound
Pivoting Factorization: A Compact Meta Low-Rank Representation of Sparsity for Efficient Inference in Large Language Models Enhancing One-shot Pruned Pre-trained Language Models through Sparse-Dense-Sparse Mechanism
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 1c687e0b-f6cc-4db9-8edf-d3db74b9303d · outbound
Pivoting Factorization: A Compact Meta Low-Rank Representation of Sparsity for Efficient Inference in Large Language Models MoDeGPT: Modular Decomposition for Large Language Model Compression
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e56ae80e-bff1-4ad7-8e15-f25de16911ae · outbound
Pivoting Factorization: A Compact Meta Low-Rank Representation of Sparsity for Efficient Inference in Large Language Models W., and Yang, Y
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation a863cb39-ff26-408d-8098-5e42facfdab0 · outbound
Pivoting Factorization: A Compact Meta Low-Rank Representation of Sparsity for Efficient Inference in Large Language Models Llm-pruner: On the structural pruning of large language models
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation f2214274-4e3b-4eb7-a56c-41b98a3dc989 · outbound
Pivoting Factorization: A Compact Meta Low-Rank Representation of Sparsity for Efficient Inference in Large Language Models Language Models are Few-Shot Learners
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cc0f972f-bc3d-456f-8eea-7c793495f5e1 · outbound
Pivoting Factorization: A Compact Meta Low-Rank Representation of Sparsity for Efficient Inference in Large Language Models ShortGPT: Layers in Large Language Models are More Redundant Than You Expect
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 37db8eef-4460-4dea-9658-bf8d49b8221e · outbound
Pivoting Factorization: A Compact Meta Low-Rank Representation of Sparsity for Efficient Inference in Large Language Models Pointer sentinel mixture models
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 72eab868-ded5-4a99-af8f-ba4d38b9a97e · outbound
Pivoting Factorization: A Compact Meta Low-Rank Representation of Sparsity for Efficient Inference in Large Language Models Accelerating Sparse Deep Neural Networks
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8e510d20-16fb-41a7-aca3-361072dab379 · outbound
Pivoting Factorization: A Compact Meta Low-Rank Representation of Sparsity for Efficient Inference in Large Language Models C., Mocanu, E., Stone, P., Nguyen, P
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 61bfff8c-5cbd-4657-bf31-85aebd8d1283 · outbound
Pivoting Factorization: A Compact Meta Low-Rank Representation of Sparsity for Efficient Inference in Large Language Models Dobi-svd: Differentiable svd for llm compression and some new perspectives
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 3ec9aa3e-05f9-41a9-8d6e-a5dfd2f45591 · outbound
Pivoting Factorization: A Compact Meta Low-Rank Representation of Sparsity for Efficient Inference in Large Language Models Improving language understanding by generative pre-training
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2b457680-39ee-4f8e-92c0-96175874730f · outbound
Pivoting Factorization: A Compact Meta Low-Rank Representation of Sparsity for Efficient Inference in Large Language Models Language models are unsupervised multitask learners
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ef7d5ff2-ec60-4d16-a14d-9ab5b3e0f031 · outbound
Pivoting Factorization: A Compact Meta Low-Rank Representation of Sparsity for Efficient Inference in Large Language Models Unresolved cited work
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation a1ddcf52-c41b-4df2-99a9-11d60a2a0df2 · outbound
Pivoting Factorization: A Compact Meta Low-Rank Representation of Sparsity for Efficient Inference in Large Language Models Compressing large language models using low rank and low precision decomposition
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 99b0df39-923e-4c01-80aa-afa92dbd32ff · outbound
Pivoting Factorization: A Compact Meta Low-Rank Representation of Sparsity for Efficient Inference in Large Language Models and Khailany, B
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 88c84fb1-8e3e-449e-b1cb-e6672d3ee367 · outbound
Pivoting Factorization: A Compact Meta Low-Rank Representation of Sparsity for Efficient Inference in Large Language Models The Truth is in There: Improving Reasoning in Language Models with Layer-Selective Rank Reduction
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ab2be93b-128f-468c-a751-fc99e8dd5c96 · outbound
Pivoting Factorization: A Compact Meta Low-Rank Representation of Sparsity for Efficient Inference in Large Language Models Sleb: streamlining llms through redundancy verification and elimination of transformer blocks
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation a56ea408-e8bb-42f5-97fa-43534bed8e5d · outbound
Pivoting Factorization: A Compact Meta Low-Rank Representation of Sparsity for Efficient Inference in Large Language Models Unresolved cited work
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 84ec4122-2bfe-403e-bd79-a6a7a0d5fba8 · outbound
Pivoting Factorization: A Compact Meta Low-Rank Representation of Sparsity for Efficient Inference in Large Language Models LLaMA: Open and Efficient Foundation Language Models
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7c36c06d-921a-4ffd-8b67-64388ae0addf · outbound
Pivoting Factorization: A Compact Meta Low-Rank Representation of Sparsity for Efficient Inference in Large Language Models Llama 2: Open Foundation and Fine-Tuned Chat Models
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8b87c97e-eb81-4e9f-aaeb-0f1324b5d3c6 · outbound
Pivoting Factorization: A Compact Meta Low-Rank Representation of Sparsity for Efficient Inference in Large Language Models Unresolved cited work
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation aace18d2-f7df-4958-aab2-d7478aaebc42 · outbound
Pivoting Factorization: A Compact Meta Low-Rank Representation of Sparsity for Efficient Inference in Large Language Models Efficient Large Language Models: A Survey
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 869380b8-aea1-4a08-a31e-f0349f26ca2b · outbound
Pivoting Factorization: A Compact Meta Low-Rank Representation of Sparsity for Efficient Inference in Large Language Models Superglue: A stickier benchmark for general-purpose language understanding systems
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 27c2739e-a650-4553-8fc6-a3035abd9333 · outbound
Pivoting Factorization: A Compact Meta Low-Rank Representation of Sparsity for Efficient Inference in Large Language Models SVD-LLM: Truncation-aware Singular Value Decomposition for Large Language Model Compression
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5311baae-38fd-4dac-8aa6-7cf4d3da9c8d · outbound
Pivoting Factorization: A Compact Meta Low-Rank Representation of Sparsity for Efficient Inference in Large Language Models Good subnetworks provably exist: Pruning via greedy forward selection
Reference 40
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 05ba53d6-5a43-4ce7-9cec-3cafa5d3360c · outbound
Pivoting Factorization: A Compact Meta Low-Rank Representation of Sparsity for Efficient Inference in Large Language Models K., Pechenizkiy, M., Liang, Y., et al
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 29897305-d38d-4d8e-b069-a3bca240c2c5 · outbound
Pivoting Factorization: A Compact Meta Low-Rank Representation of Sparsity for Efficient Inference in Large Language Models ASVD: Activation-aware Singular Value Decomposition for Compressing Large Language Models
Reference 42
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 42e23edb-2199-44c2-847d-59e333cbbebb · outbound
Pivoting Factorization: A Compact Meta Low-Rank Representation of Sparsity for Efficient Inference in Large Language Models Unresolved cited work
Reference 43
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation bcaffdae-733e-481d-a786-09b38cdee119 · outbound
Pivoting Factorization: A Compact Meta Low-Rank Representation of Sparsity for Efficient Inference in Large Language Models To prune, or not to prune: exploring the efficacy of pruning for model compression
Reference 44
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 56c7405d-5a29-4a4c-93fc-ff5af877dbb5 · outbound
Pivoting Factorization: A Compact Meta Low-Rank Representation of Sparsity for Efficient Inference in Large Language Models A survey on model compression for large language models
Reference 45
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7201c720-63f7-41cc-ad86-e3b0cd1af598 · outbound
Pivoting Factorization: A Compact Meta Low-Rank Representation of Sparsity for Efficient Inference in Large Language Models Discrimination-aware channel pruning for deep neural networks
Reference 46
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 267a814a-ca6f-4284-aece-018a7fad473f · outbound
Pivoting Factorization: A Compact Meta Low-Rank Representation of Sparsity for Efficient Inference in Large Language Models write newline
Reference 47
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0f894493-e622-4418-a1e9-097317145de2 · inbound
Accelerating Attention with Basis Decomposition Pivoting Factorization: A Compact Meta Low-Rank Representation of Sparsity for Efficient Inference in Large Language Models
Reference 60
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b7ff421f-9433-4a1e-9a84-84ec82db55f1 · inbound
Understanding Calibration and Truncation Error Propagation in Training-Free Low-Rank Compression for LLMs Pivoting Factorization: A Compact Meta Low-Rank Representation of Sparsity for Efficient Inference in Large Language Models
Reference 86
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.