Pith. sign in

Paper Citation Record · LEDGER

Outlier Suppression+: Accurate quantization of large language models by equivalent and optimal shifting and scaling

As of 14 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 29 inbound Pith citation observations for arXiv:2304.09145.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2304.09145 v3

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 29 of 29 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-14T06:32:32.682623+00:00

measured 29 of 29 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-14T12:54:26.620007Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-06-28T19:42:36.070011Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation e73f5471-146c-465c-8514-1674f4582101 · inbound

A Comprehensive Overview of Large Language Models cites this paper.

A Comprehensive Overview of Large Language Models Outlier Suppression+: Accurate quantization of large language models by equivalent and optimal shifting and scaling

Reference 257

Resolution
verified exact
arxiv_id, observed 2026-05-19T20:28:39.159884Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-05-19T20:28:38.900026Z digest=sha256:29c2c24a3e3e8eb726af878aa7293fcadc6e199735c82279fec58492b4c131dc

Observation 34588567-aa6b-4b47-ba76-f31aeac5d949 · inbound

A Survey on Efficient Inference for Large Language Models cites this paper.

A Survey on Efficient Inference for Large Language Models Outlier Suppression+: Accurate quantization of large language models by equivalent and optimal shifting and scaling

Reference 169

Resolution
verified exact
arxiv_id, observed 2026-05-15T02:39:33.108405Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-05-15T02:39:33.007894Z digest=sha256:4d224f99dce3fb241751de20f86a540263dffca60324de3dfdf3981701e5792a

Observation 0c85add7-2e8a-4b72-b67c-98d6ba5d6a7f · inbound

SpinQuant: LLM quantization with learned rotations cites this paper.

SpinQuant: LLM quantization with learned rotations Outlier Suppression+: Accurate quantization of large language models by equivalent and optimal shifting and scaling

Reference 22

Resolution
verified exact
arxiv_id, observed 2026-05-15T15:52:34.788483Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-05-15T15:52:34.606853Z digest=sha256:da7af871f3703201e4d4317f2c286f16e4f1e20c2a1ba4c09240c27f7f81e1f1

Observation bd7249e2-90cd-4f89-b4ac-be228cefbd4c · inbound

DFRot: Achieving Outlier-Free and Massive Activation-Free for Rotated LLMs with Refined Rotation cites this paper.

DFRot: Achieving Outlier-Free and Massive Activation-Free for Rotated LLMs with Refined Rotation Outlier Suppression+: Accurate quantization of large language models by equivalent and optimal shifting and scaling

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-12T05:20:23.935806Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T05:20:23.935806Z digest=sha256:6c72f1a758186deb91fded9799f596fe5adf8da8acad3782a5ca9665dc044f54

Observation 288e1c28-d624-48ca-97fe-73cdf11e8452 · inbound

Deploying Foundation Model Powered Agent Services: A Survey cites this paper.

Deploying Foundation Model Powered Agent Services: A Survey Outlier Suppression+: Accurate quantization of large language models by equivalent and optimal shifting and scaling

Reference 200

Resolution
unresolved
no resolver link, observed 2026-08-11T13:09:46.627672Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T13:09:46.627672Z digest=sha256:5c9e67367b36475e5814ba3abc269a90543988b5f12e766297ff4efa2021299a

Observation 3187b729-b6b5-4683-934a-b59df0dea78e · inbound

A Survey on Large Language Model Acceleration based on KV Cache Management cites this paper.

A Survey on Large Language Model Acceleration based on KV Cache Management Outlier Suppression+: Accurate quantization of large language models by equivalent and optimal shifting and scaling

Reference 82

Resolution
unresolved
no resolver link, observed 2026-08-11T00:38:47.273784Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T00:38:47.273784Z digest=sha256:fde78fc2815ad829b2b416bc86fec95d537b836f4a312ac55c36890ab8e2297d

Observation 1f82671e-6d1e-403d-aee3-7da93505fcfe · inbound

Highly Optimized Kernels and Fine-Grained Codebooks for LLM Inference on Arm CPUs cites this paper.

Highly Optimized Kernels and Fine-Grained Codebooks for LLM Inference on Arm CPUs Outlier Suppression+: Accurate quantization of large language models by equivalent and optimal shifting and scaling

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-11T05:45:50.287457Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T05:45:50.287457Z digest=sha256:fdd82b852413524a7975e9a633fb31746406623e61adabb5ae968fa7b084b909

Observation 055707ab-0c07-45ed-b1f4-75872dfe7a32 · inbound

Dissecting Bit-Level Scaling Laws in Quantizing Vision Generative Models cites this paper.

Dissecting Bit-Level Scaling Laws in Quantizing Vision Generative Models Outlier Suppression+: Accurate quantization of large language models by equivalent and optimal shifting and scaling

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-10T22:04:24.319774Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T22:04:24.319774Z digest=sha256:020df5ce8f6f72b0e0e3a36aeb7e71d93b8ed9bd84c3f99c6d56dbbabe14cfea

Observation c208a281-ae88-4a78-8607-fa24e89187bb · inbound

Accelerating Large Language Models through Partially Linear Feed-Forward Network cites this paper.

Accelerating Large Language Models through Partially Linear Feed-Forward Network Outlier Suppression+: Accurate quantization of large language models by equivalent and optimal shifting and scaling

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-10T19:29:57.469890Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T19:29:57.469890Z digest=sha256:637e5d6d6ae9f9fca4373594f65ec8f7507642093b9ec44f410f383081723a24

Observation 88b2d013-29a9-4bea-b67a-e56b73f4af4a · inbound

Qrazor: Reliable and Effortless 4-bit LLM Quantization by Significant Data Razoring cites this paper.

Qrazor: Reliable and Effortless 4-bit LLM Quantization by Significant Data Razoring Outlier Suppression+: Accurate quantization of large language models by equivalent and optimal shifting and scaling

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-10T16:19:57.610432Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T16:19:57.610432Z digest=sha256:1d21d7956b7b03c976fc411288eeeba65fc08969f698639d3e81263fa9bdfc25

Observation fa6a0ec5-f35c-433f-9f9d-556e71043bd8 · inbound

On Accelerating Edge AI: Optimizing Resource-Constrained Environments cites this paper.

On Accelerating Edge AI: Optimizing Resource-Constrained Environments Outlier Suppression+: Accurate quantization of large language models by equivalent and optimal shifting and scaling

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-10T14:46:38.167424Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T14:46:38.167424Z digest=sha256:6fb9122c3da035565af1bb50226c7156687d25fd0fb3bff6b8d3e5892416691c

Observation 9ed1d098-b6a5-42f7-814f-85f2d927af3d · inbound

Quaff: Quantized Parameter-Efficient Fine-Tuning under Outlier Spatial Stability Hypothesis cites this paper.

Quaff: Quantized Parameter-Efficient Fine-Tuning under Outlier Spatial Stability Hypothesis Outlier Suppression+: Accurate quantization of large language models by equivalent and optimal shifting and scaling

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-07T15:43:41.625021Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:43:41.625021Z digest=sha256:f11032cfca9ab4bab81463cb4eddbe6c22127d8329685c1f09c2447081c8ed52

Observation 633f0dc7-212f-47b7-864f-b5b08e2e0762 · inbound

NQKV: A KV Cache Quantization Scheme Based on Normal Distribution Characteristics cites this paper.

NQKV: A KV Cache Quantization Scheme Based on Normal Distribution Characteristics Outlier Suppression+: Accurate quantization of large language models by equivalent and optimal shifting and scaling

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-07T15:09:28.972282Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:09:28.972282Z digest=sha256:e5db1424204e5ccd8ca91976ab1cf62f850254f91187f5a1fe423994bb418bc2

Observation b7a28310-e77c-47cf-a0e2-19d3232e669b · inbound

KerZOO: Kernel Function Informed Zeroth-Order Optimization for Accurate and Accelerated LLM Fine-Tuning cites this paper.

KerZOO: Kernel Function Informed Zeroth-Order Optimization for Accurate and Accelerated LLM Fine-Tuning Outlier Suppression+: Accurate quantization of large language models by equivalent and optimal shifting and scaling

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-07T14:29:34.477636Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:29:34.477636Z digest=sha256:9e4edc3500bda05050c481f1769d6f6c1bfe00b10deaee9b0e87d09296a68873

Observation 4ae87194-27b8-4ff2-b320-de5ba10837d1 · inbound

DECA: A Near-Core LLM Decompression Accelerator Grounded on a 3D Roofline Model cites this paper.

DECA: A Near-Core LLM Decompression Accelerator Grounded on a 3D Roofline Model Outlier Suppression+: Accurate quantization of large language models by equivalent and optimal shifting and scaling

Reference 75

Resolution
unresolved
no resolver link, observed 2026-08-07T14:21:58.025701Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:21:58.025701Z digest=sha256:89d0b5fa40e806e18dd5ded0ac6c187be4bbd52f4d09c0d61e95cd654c34571a

Observation 0af6810a-da2d-419c-840e-3bac0083962f · inbound

Rethinking the Outlier Distribution in Large Language Models: An In-depth Study cites this paper.

Rethinking the Outlier Distribution in Large Language Models: An In-depth Study Outlier Suppression+: Accurate quantization of large language models by equivalent and optimal shifting and scaling

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-07T13:30:40.860196Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:30:40.860196Z digest=sha256:43e80908a8b33b5e589428b2f400218d6410f7cb597ca1f12f6ca50230722fb1

Observation a2cac9bb-05ea-4ff6-a91f-d0c1ecd524df · inbound

FPTQuant: Function-Preserving Transforms for LLM Quantization cites this paper.

FPTQuant: Function-Preserving Transforms for LLM Quantization Outlier Suppression+: Accurate quantization of large language models by equivalent and optimal shifting and scaling

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-07T10:43:46.980855Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T10:43:46.980855Z digest=sha256:dc2cef9aca85d87d1039a0813b42489caaddb6862a22ce437328cb2206d89794

Observation a05345fa-7475-41b2-a502-2699e1cf9580 · inbound

BASE-Q: Bias and Asymmetric Scaling Enhanced Rotational Quantization for Large Language Models cites this paper.

BASE-Q: Bias and Asymmetric Scaling Enhanced Rotational Quantization for Large Language Models Outlier Suppression+: Accurate quantization of large language models by equivalent and optimal shifting and scaling

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-07T14:06:24.060629Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:06:24.060629Z digest=sha256:0dca7ef1f20deb7e8c241ac9e4ccd39b9b898d2eaae8c2ad2f133a86a6c4e63b

Observation 32bf6349-d0e6-4e50-bce4-50b815969fc6 · inbound

Q-resafe: Assessing Safety Risks and Quantization-aware Safety Patching for Quantized Large Language Models cites this paper.

Q-resafe: Assessing Safety Risks and Quantization-aware Safety Patching for Quantized Large Language Models Outlier Suppression+: Accurate quantization of large language models by equivalent and optimal shifting and scaling

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-06T23:00:19.611331Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:00:19.611331Z digest=sha256:aa58db7e4672d360c8ee6629125c1d43c6207f3a8ea89937dbc49048c68d1227

Observation f247830d-ff20-4990-83eb-6e8bebe5da40 · inbound

PoTPTQ: A Two-step Power-of-Two Post-training for LLMs cites this paper.

PoTPTQ: A Two-step Power-of-Two Post-training for LLMs Outlier Suppression+: Accurate quantization of large language models by equivalent and optimal shifting and scaling

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-06T17:06:36.149236Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:06:36.149236Z digest=sha256:d31141c2c1b875816628f7945a840a18dbdd5c6e1ebcee36d39b887aae3bc987

Observation 6b86ac59-2285-4dfb-8993-1e5bd2975a13 · inbound

DMQ: Dissecting Outliers of Diffusion Models for Post-Training Quantization cites this paper.

DMQ: Dissecting Outliers of Diffusion Models for Post-Training Quantization Outlier Suppression+: Accurate quantization of large language models by equivalent and optimal shifting and scaling

Reference 63

Resolution
unresolved
no resolver link, observed 2026-08-06T16:42:12.194747Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:42:12.194747Z digest=sha256:83ad59ea49322b9733785af5b18090a28b638193f40f57e87bf1344b84b8cd80

Observation 057bd4f2-a8c0-4035-9284-998b1ad53b9a · inbound

LiquidGEMM: Hardware-Efficient W4A8 GEMM Kernel for High-Performance LLM Serving cites this paper.

LiquidGEMM: Hardware-Efficient W4A8 GEMM Kernel for High-Performance LLM Serving Outlier Suppression+: Accurate quantization of large language models by equivalent and optimal shifting and scaling

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-05T12:51:01.202661Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T12:51:01.202661Z digest=sha256:645330bd0bfe6189a0578f148eb4444e8401059c4185bf18904367ab5ef98ad8

Observation 93c423a1-59f0-4a34-9066-41f91a54c11b · inbound

Efficient Reasoning on the Edge cites this paper.

Efficient Reasoning on the Edge Outlier Suppression+: Accurate quantization of large language models by equivalent and optimal shifting and scaling

Reference 135

Resolution
unresolved
no resolver link, observed 2026-07-13T23:28:12.790404Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-13T23:28:12.790404Z digest=sha256:9859ef8d23a875e4f8691de6714f1440adc0280ba918ee33516547fe73df7e8a

Observation cc8739a6-9c27-4149-af0d-a97be86a3177 · inbound

LBLLM: Lightweight Binarization of Large Language Models via Three-Stage Distillation cites this paper.

LBLLM: Lightweight Binarization of Large Language Models via Three-Stage Distillation Outlier Suppression+: Accurate quantization of large language models by equivalent and optimal shifting and scaling

Reference 29

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T12:46:04.675725Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=arxiv_source observed=2026-05-10T03:04:14.900791Z digest=sha256:5f070bc0d857fca2e0abab1c2d4e5c49b6c2cad469cb3970b59a914db7b24d02

Observation 72e8ee89-4cd9-434b-9197-301b6eeb3b86 · inbound

Network Edge Inference for Large Language Models: Principles, Techniques, and Opportunities cites this paper.

Network Edge Inference for Large Language Models: Principles, Techniques, and Opportunities Outlier Suppression+: Accurate quantization of large language models by equivalent and optimal shifting and scaling

Reference 163

Resolution
verified exact
arxiv_id, observed 2026-05-11T20:16:10.297825Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-05-08T09:45:57.201837Z digest=sha256:564a4dfce134c74f3130e31a6e5ddb0b04b406acef49d2317b298c9332256f7e

Observation fdf57dfa-0585-4c9d-ad2a-33e054ae9f25 · inbound

Quant.npu: Enabling Efficient Mobile NPU Inference for on-device LLMs via Fully Static Quantization cites this paper.

Quant.npu: Enabling Efficient Mobile NPU Inference for on-device LLMs via Fully Static Quantization Outlier Suppression+: Accurate quantization of large language models by equivalent and optimal shifting and scaling

Reference 38

Resolution
verified exact
arxiv_id, observed 2026-05-21T07:54:02.625596Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-05-21T07:53:48.946110Z digest=sha256:a6e39461db9e85cebc8189ef0ede87c7a667e449059c44f8a231289f0dc2cdca

Observation cd16a273-dd71-4d18-9bd5-87328511b638 · inbound

GNMR: Runtime Stability Control for Low-Precision Large Language Model Training cites this paper.

GNMR: Runtime Stability Control for Low-Precision Large Language Model Training Outlier Suppression+: Accurate quantization of large language models by equivalent and optimal shifting and scaling

Reference 57

Resolution
verified exact
arxiv_id, observed 2026-06-28T19:42:36.071574Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-06-28T18:53:47.187437Z digest=sha256:a87d085022b89916264919305a56f7caad7bfdadadc1dd723f702c4a021341ef

Observation e210193d-bab0-49e1-b61c-e1d60621ce0e · inbound

MXSens: Sensitivity-Aware Mixed-Precision Quantization for Efficient LLM Inference cites this paper.

MXSens: Sensitivity-Aware Mixed-Precision Quantization for Efficient LLM Inference Outlier Suppression+: Accurate quantization of large language models by equivalent and optimal shifting and scaling

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-01T17:12:27.226905Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T17:12:27.226905Z digest=sha256:d193df0b67b9d0470fa9fa32e116111a0e3dfc371677e4cb3adf261b2afb86f7

Observation 56f04a3c-c28b-420f-8654-1eea4d2fb4de · inbound

When Local Variance Optimality Is Not Enough: RoPE-Aligned Q/K Rotations for Dynamic 4-Bit Quantisation cites this paper.

When Local Variance Optimality Is Not Enough: RoPE-Aligned Q/K Rotations for Dynamic 4-Bit Quantisation Outlier Suppression+: Accurate quantization of large language models by equivalent and optimal shifting and scaling

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-14T12:54:26.620007Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T12:54:26.620007Z digest=sha256:2edda32a7e65737759d18e18818a70d116db29e634138230662af10ae13fe477