Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-10T16:13:20.004756Z
Paper Citation Record · LEDGER
As of 17 August 2026, this Paper Citation Record lists 37 of 37 outbound references and 22 inbound Pith citation observations for arXiv:2501.13987.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-10T16:13:20.004756Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-16T06:30:59.297886+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-14T12:54:26.511410Z
A source-named dated measurement, never combined with another source.
Source: pith, observed 2026-07-10T14:47:14.541739Z
37 of 37 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation f9e28a0a-43f5-4102-b60c-d2bf42b495e4 · outbound
OstQuant: Refining Large Language Model Quantization with Orthogonal and Scaling Transformations for Better Distribution Fitting GPT-4 Technical Report
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d96730dd-3b9b-4b4e-8f45-535fc4726c1a · outbound
OstQuant: Refining Large Language Model Quantization with Orthogonal and Scaling Transformations for Better Distribution Fitting QuaRot: Outlier-Free 4-Bit Inference in Rotated LLMs
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5a7f3044-ddba-40da-a6d2-5f8d7a476db1 · outbound
OstQuant: Refining Large Language Model Quantization with Orthogonal and Scaling Transformations for Better Distribution Fitting Riemannian Adaptive Optimization Methods
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1a1eb707-f59b-4400-890e-dc51fa180a0a · outbound
OstQuant: Refining Large Language Model Quantization with Orthogonal and Scaling Transformations for Better Distribution Fitting Piqa: Reasoning about physical commonsense in natural language
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 87526d87-580b-44a3-ac45-b82f1efe4d3b · outbound
OstQuant: Refining Large Language Model Quantization with Orthogonal and Scaling Transformations for Better Distribution Fitting A Systematic Classification of Knowledge, Reasoning, and Context within the ARC Dataset
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 50759f5d-a546-42c4-888f-72e3fee4cbcc · outbound
OstQuant: Refining Large Language Model Quantization with Orthogonal and Scaling Transformations for Better Distribution Fitting QuIP: 2-Bit Quantization of Large Language Models With Guarantees
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b983fb3d-502e-4f32-a4e8-cbd9a9de49d8 · outbound
OstQuant: Refining Large Language Model Quantization with Orthogonal and Scaling Transformations for Better Distribution Fitting Driving with llms: Fusing object-level vector modality for explainable autonomous driving
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 8aa98808-749c-480d-a83e-082e4fe38242 · outbound
OstQuant: Refining Large Language Model Quantization with Orthogonal and Scaling Transformations for Better Distribution Fitting Distributional quantization of large language models
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 2e74fef4-d385-43d9-8b0c-1f008c0f3e55 · outbound
OstQuant: Refining Large Language Model Quantization with Orthogonal and Scaling Transformations for Better Distribution Fitting BoolQ: Exploring the Surprising Difficulty of Natural Yes/No Questions
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4a5ba079-b1c3-4606-816b-65b18b35b9a4 · outbound
OstQuant: Refining Large Language Model Quantization with Orthogonal and Scaling Transformations for Better Distribution Fitting Unresolved cited work
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a0be1eaa-e935-424c-b151-c8778f59b39d · outbound
OstQuant: Refining Large Language Model Quantization with Orthogonal and Scaling Transformations for Better Distribution Fitting GPTQ: Accurate Post-Training Quantization for Generative Pre-trained Transformers
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d158af50-efbe-404e-a5b5-dd08d3199988 · outbound
OstQuant: Refining Large Language Model Quantization with Orthogonal and Scaling Transformations for Better Distribution Fitting A framework for few-shot language model evaluation, 07 2024
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation aaaa6f37-4691-448f-b13b-563a8c370dec · outbound
OstQuant: Refining Large Language Model Quantization with Orthogonal and Scaling Transformations for Better Distribution Fitting I-LLM: Efficient Integer-Only Inference for Fully-Quantized Low-Bit Large Language Models
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 173008cf-539b-4b48-bf75-2b117e72f090 · outbound
OstQuant: Refining Large Language Model Quantization with Orthogonal and Scaling Transformations for Better Distribution Fitting The topology of Stiefel manifolds, volume 24
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation ecc961d3-0e04-44d8-9d9b-79cb1e4398bc · outbound
OstQuant: Refining Large Language Model Quantization with Orthogonal and Scaling Transformations for Better Distribution Fitting Adam: A Method for Stochastic Optimization
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 57493104-f9ce-4cd2-a527-6bba40b14ace · outbound
OstQuant: Refining Large Language Model Quantization with Orthogonal and Scaling Transformations for Better Distribution Fitting Geoopt: Riemannian optimization in pytorch, 2020
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8099866e-fca7-4610-b48f-6d02622d2c7e · outbound
OstQuant: Refining Large Language Model Quantization with Orthogonal and Scaling Transformations for Better Distribution Fitting OWQ: Outlier-Aware Weight Quantization for Efficient Fine-Tuning and Inference of Large Language Models
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0f5b7e03-ee08-4fc2-afb7-9614d5a8acaa · outbound
OstQuant: Refining Large Language Model Quantization with Orthogonal and Scaling Transformations for Better Distribution Fitting Efficient Riemannian Optimization on the Stiefel Manifold via the Cayley Transform
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 930d2d93-5223-4a70-aad2-3647360bb88c · outbound
OstQuant: Refining Large Language Model Quantization with Orthogonal and Scaling Transformations for Better Distribution Fitting AWQ: Activation-aware Weight Quantization for LLM Compression and Acceleration
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9a126cef-7a74-45cb-aa0b-3c97758022b9 · outbound
OstQuant: Refining Large Language Model Quantization with Orthogonal and Scaling Transformations for Better Distribution Fitting SpinQuant: LLM quantization with learned rotations
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 166a56b1-ef19-463e-839d-1cb62672ca48 · outbound
OstQuant: Refining Large Language Model Quantization with Orthogonal and Scaling Transformations for Better Distribution Fitting AffineQuant: Affine Transformation Quantization for Large Language Models
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6a645625-2a43-4db3-818a-4bdb75671c5b · outbound
OstQuant: Refining Large Language Model Quantization with Orthogonal and Scaling Transformations for Better Distribution Fitting Pointer Sentinel Mixture Models
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 01b1de57-5097-44b0-b57e-c2821d653482 · outbound
OstQuant: Refining Large Language Model Quantization with Orthogonal and Scaling Transformations for Better Distribution Fitting Can a Suit of Armor Conduct Electricity? A New Dataset for Open Book Question Answering
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7c0038c3-37cd-4fd9-bc2c-522ec5ac7ffc · outbound
OstQuant: Refining Large Language Model Quantization with Orthogonal and Scaling Transformations for Better Distribution Fitting Language models are unsupervised multitask learners
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1fd6c393-d5d7-46df-b17d-270372abe28d · outbound
OstQuant: Refining Large Language Model Quantization with Orthogonal and Scaling Transformations for Better Distribution Fitting Winogrande: An adversarial winograd schema challenge at scale
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 67d41117-7f90-49cc-8211-4eeed2a23e34 · outbound
OstQuant: Refining Large Language Model Quantization with Orthogonal and Scaling Transformations for Better Distribution Fitting SocialIQA: Commonsense Reasoning about Social Interactions
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation df928ba2-f7dc-455a-b9b7-4c19c81a4607 · outbound
OstQuant: Refining Large Language Model Quantization with Orthogonal and Scaling Transformations for Better Distribution Fitting OmniQuant: Omnidirectionally Calibrated Quantization for Large Language Models
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 90b7234d-c566-4752-880b-594c5e760a1d · outbound
OstQuant: Refining Large Language Model Quantization with Orthogonal and Scaling Transformations for Better Distribution Fitting LLaMA: Open and Efficient Foundation Language Models
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 71a20098-ff60-41a4-bd5f-56be22d2b7a0 · outbound
OstQuant: Refining Large Language Model Quantization with Orthogonal and Scaling Transformations for Better Distribution Fitting Llama 2: Open Foundation and Fine-Tuned Chat Models
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6537f63c-1669-4b9b-a292-4bb7a9402611 · outbound
OstQuant: Refining Large Language Model Quantization with Orthogonal and Scaling Transformations for Better Distribution Fitting QuIP#: Even Better LLM Quantization with Hadamard Incoherence and Lattice Codebooks
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a8e874f3-822b-48d7-af15-c9534422461c · outbound
OstQuant: Refining Large Language Model Quantization with Orthogonal and Scaling Transformations for Better Distribution Fitting SmoothQuant: Accurate and Efficient Post-Training Quantization for Large Language Models
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d1c187c5-2f9a-4c8e-b164-77642734c8de · outbound
OstQuant: Refining Large Language Model Quantization with Orthogonal and Scaling Transformations for Better Distribution Fitting ZeroQuant: Efficient and Affordable Post-Training Quantization for Large-Scale Transformers
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1e4030b3-c143-4615-9710-868a7e384154 · outbound
OstQuant: Refining Large Language Model Quantization with Orthogonal and Scaling Transformations for Better Distribution Fitting HellaSwag: Can a Machine Really Finish Your Sentence?
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ff7a0792-c45d-43dd-beb7-3ba3753d3629 · outbound
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 68bf4dff-eb60-4194-876c-414bde825a80 · outbound
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 853873af-147e-4eea-bcc5-073c0c79443f · outbound
OstQuant: Refining Large Language Model Quantization with Orthogonal and Scaling Transformations for Better Distribution Fitting Unresolved cited work
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 72d7de02-d8d0-49d8-94a9-0dbc2585bd05 · outbound
OstQuant: Refining Large Language Model Quantization with Orthogonal and Scaling Transformations for Better Distribution Fitting Unresolved cited work
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5fe2956b-78de-4d3e-ba8c-ce8102e30ea0 · inbound
DFRot: Achieving Outlier-Free and Massive Activation-Free for Rotated LLMs with Refined Rotation OstQuant: Refining Large Language Model Quantization with Orthogonal and Scaling Transformations for Better Distribution Fitting
Reference 2020
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a559fc8b-f7ee-40ab-b741-7b5fc3d2db5f · inbound
MQuant: Unleashing the Inference Potential of Multimodal Large Language Models via Full Static Quantization OstQuant: Refining Large Language Model Quantization with Orthogonal and Scaling Transformations for Better Distribution Fitting
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 76dddf82-afb5-4854-bccc-a09f17123e90 · inbound
NeUQI: Near-Optimal Uniform Quantization Parameter Initialization for Low-Bit LLMs OstQuant: Refining Large Language Model Quantization with Orthogonal and Scaling Transformations for Better Distribution Fitting
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 67e9cea5-60b0-4f83-a1c7-19a6a6315861 · inbound
NSNQuant: A Double Normalization Approach for Calibration-Free Low-Bit Vector Quantization of KV Cache OstQuant: Refining Large Language Model Quantization with Orthogonal and Scaling Transformations for Better Distribution Fitting
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d79a8ae5-d216-42f4-a1ae-10e1599b4b39 · inbound
TAH-QUANT: Effective Activation Quantization in Pipeline Parallelism over Slow Network OstQuant: Refining Large Language Model Quantization with Orthogonal and Scaling Transformations for Better Distribution Fitting
Reference 43
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 4f427d71-a399-4caf-ab73-9d9024678da8 · inbound
PCDVQ: Enhancing Vector Quantization for Large Language Models via Polar Coordinate Decoupling OstQuant: Refining Large Language Model Quantization with Orthogonal and Scaling Transformations for Better Distribution Fitting
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9c216eaf-d8bd-49af-87fe-7753206914e9 · inbound
BTC-LLM: Efficient Sub-1-Bit LLM Quantization via Learnable Transformation and Binary Codebook OstQuant: Refining Large Language Model Quantization with Orthogonal and Scaling Transformations for Better Distribution Fitting
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 6d56403d-f536-4b30-b255-95fc7fe68239 · inbound
BASE-Q: Bias and Asymmetric Scaling Enhanced Rotational Quantization for Large Language Models OstQuant: Refining Large Language Model Quantization with Orthogonal and Scaling Transformations for Better Distribution Fitting
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 332b1750-a916-408a-9724-90631d4a2c20 · inbound
QuantVLA: Scale-Calibrated Post-Training Quantization for Vision-Language-Action Models OstQuant: Refining Large Language Model Quantization with Orthogonal and Scaling Transformations for Better Distribution Fitting
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 7a238a67-7853-49d2-9596-37552a95dd6b · inbound
CoQuant: Joint Weight-Activation Subspace Projection for Mixed-Precision LLMs OstQuant: Refining Large Language Model Quantization with Orthogonal and Scaling Transformations for Better Distribution Fitting
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 9d909944-23ac-45db-bc4c-45acc74a9ede · inbound
LoopQ: Quantization for Recursive Transformers OstQuant: Refining Large Language Model Quantization with Orthogonal and Scaling Transformations for Better Distribution Fitting
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 1f2f61f0-e9d5-4cf8-9ad8-f9a6ec79edf2 · inbound
GAMMA: Global Bit Allocation for Mixed-Precision Models under Arbitrary Budgets OstQuant: Refining Large Language Model Quantization with Orthogonal and Scaling Transformations for Better Distribution Fitting
Reference 49
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation f79f4af6-d087-43fc-9b2f-f646131dfb32 · inbound
Theory-optimal Quantization Based on Flatness OstQuant: Refining Large Language Model Quantization with Orthogonal and Scaling Transformations for Better Distribution Fitting
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation b600cc3a-2a90-4409-8fc9-e66fa77b68e6 · inbound
Breaking Modality Heterogeneity in Low-Bit Quantization for Large Vision-Language Models OstQuant: Refining Large Language Model Quantization with Orthogonal and Scaling Transformations for Better Distribution Fitting
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation af2a470a-7516-4cd6-93cf-417adf65cc40 · inbound
MGVQ: Synergizing Multi-dimensional Sensitivity-Aware and Gradient-Hessian Fusion for Vector Quantization OstQuant: Refining Large Language Model Quantization with Orthogonal and Scaling Transformations for Better Distribution Fitting
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 6a58618d-e155-46f8-80b3-7768f06d4160 · inbound
HoloQ-VLA: Uniform W4A4 Quantization of Vision-Language-Action Models OstQuant: Refining Large Language Model Quantization with Orthogonal and Scaling Transformations for Better Distribution Fitting
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 83fea8c0-1853-41f1-94ad-521d2a547e1a · inbound
MixFP4: Enhancing NVFP4 with Adaptive FP4/INT4 Block Representations OstQuant: Refining Large Language Model Quantization with Orthogonal and Scaling Transformations for Better Distribution Fitting
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 7542aa95-3ac1-4b1c-8cea-3e8ba3d8b8e6 · inbound
OrbitQuant: Data-Agnostic Quantization for Image and Video Diffusion Transformers OstQuant: Refining Large Language Model Quantization with Orthogonal and Scaling Transformations for Better Distribution Fitting
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation eb0e06c6-54a2-4af6-843d-8aee223ec01b · inbound
KronQ: LLM Quantization via Kronecker-Factored Hessian OstQuant: Refining Large Language Model Quantization with Orthogonal and Scaling Transformations for Better Distribution Fitting
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation ba66d4ad-6293-4e9e-ad0b-b288e8cebc40 · inbound
Break Through the Compression Bottleneck: From Theory to Practice OstQuant: Refining Large Language Model Quantization with Orthogonal and Scaling Transformations for Better Distribution Fitting
Reference 81
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a0da8510-6be4-4114-a83f-c86f6c1a7264 · inbound
GyRot: Leveraging Hidden Synergy between Rotation and Fine-grained Group Quantization for Low-bit LLM Inference OstQuant: Refining Large Language Model Quantization with Orthogonal and Scaling Transformations for Better Distribution Fitting
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ab679e37-2974-4fad-ac33-4266e21ecace · inbound
When Local Variance Optimality Is Not Enough: RoPE-Aligned Q/K Rotations for Dynamic 4-Bit Quantisation OstQuant: Refining Large Language Model Quantization with Orthogonal and Scaling Transformations for Better Distribution Fitting
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.