Pith. sign in

Paper Citation Record · LEDGER

Reducing Storage of Pretrained Neural Networks by Rate-Constrained Quantization and Entropy Coding

As of 8 August 2026, this Paper Citation Record lists 58 of 58 outbound references and 0 inbound Pith citation observations for arXiv:2505.18758.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.18758 v1

Coverage vector

measured 58 of 58 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T14:32:20.014194Z

measured 58 of 58 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

58 of 58 outbound references displayed

  • verified exact0
  • verified fuzzy51
  • unresolved7
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 40af015c-b43e-4cbc-81e1-a8cdcc552ddf · outbound

This paper cites Croci, Bo Li, Pashmina Cameron, Martin Jaggi, Dan Alistarh, Torsten Hoefler, and James Hensman.

Reducing Storage of Pretrained Neural Networks by Rate-Constrained Quantization and Entropy Coding Croci, Bo Li, Pashmina Cameron, Martin Jaggi, Dan Alistarh, Torsten Hoefler, and James Hensman

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:32:28.797709Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:32:16.190894Z digest=sha256:7297a98dfff86bf814fc14b3cdc36621bf0f40921d0fe811764252d06220f0ae

Observation c51b9444-a97c-4766-a862-6c45b7db9d9c · outbound

This paper cites GPTVQ: The blessing of dimensionality for LLM quantization.

Reducing Storage of Pretrained Neural Networks by Rate-Constrained Quantization and Entropy Coding GPTVQ: The blessing of dimensionality for LLM quantization

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:32:28.654010Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:32:16.241300Z digest=sha256:926ea80d626d57e721852b47e743efa77efb713433095042076eff75bad689ae

Observation e4eba4ad-b0b8-47bc-a21b-5627f02ccf0b · outbound

This paper cites ONNX: Open neural network exchange, 2019.

Reducing Storage of Pretrained Neural Networks by Rate-Constrained Quantization and Entropy Coding ONNX: Open neural network exchange, 2019

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:32:28.545358Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:32:16.358834Z digest=sha256:508b0a02686eb88e083c48df552f5574467e454d04ac8f05b4a70e76315bcecf

Observation b84cb3ec-0010-460b-a007-edfffe9f8604 · outbound

This paper cites Understanding Entropy Coding With Asymmetric Numeral Systems (ANS): a Statistician's Perspective.

Reducing Storage of Pretrained Neural Networks by Rate-Constrained Quantization and Entropy Coding Understanding Entropy Coding With Asymmetric Numeral Systems (ANS): a Statistician's Perspective

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-07T14:32:16.426671Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:32:16.426671Z digest=sha256:ace5500157294555a705021f2d1fbab1d8fe941348559cf157c4cad815e299c4

Observation c74d72d4-3de6-4958-b936-3e752748c3cb · outbound

This paper cites Bronstein, and Avi Mendelson.

Reducing Storage of Pretrained Neural Networks by Rate-Constrained Quantization and Entropy Coding Bronstein, and Avi Mendelson

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:32:28.368746Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:32:16.511794Z digest=sha256:ad7064873f2ea7526b33c6427e0ebf5fb180d04db6df50932d8096bc33a85048

Observation dc197d84-ff2f-4ee5-8f6e-447a5cd3dde0 · outbound

This paper cites NNCodec: An Open Source Software Implementation of the Neural Network Coding ISO/IEC Standard.

Reducing Storage of Pretrained Neural Networks by Rate-Constrained Quantization and Entropy Coding NNCodec: An Open Source Software Implementation of the Neural Network Coding ISO/IEC Standard

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:32:28.181108Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:32:16.583499Z digest=sha256:6611dc5348f6ac7332f83566a01223da61692335451e981c61421acd5e9a8b4e

Observation fcc570e4-20a8-4a20-ac16-835174b597f0 · outbound

This paper cites EfficientQAT: Efficient Quantization-Aware Training for Large Language Models, October 2024.

Reducing Storage of Pretrained Neural Networks by Rate-Constrained Quantization and Entropy Coding EfficientQAT: Efficient Quantization-Aware Training for Large Language Models, October 2024

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:32:27.999615Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:32:16.640320Z digest=sha256:5de7d440c9a12e000b61c3d39343f97e16df632acf5c91b4d42d2a9fe9fd9a9c

Observation 98f4c2c9-2ba9-42fb-823e-77ad995f07db · outbound

This paper cites Bronstein, and Avi Mendelson.

Reducing Storage of Pretrained Neural Networks by Rate-Constrained Quantization and Entropy Coding Bronstein, and Avi Mendelson

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:32:27.831968Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:32:16.729492Z digest=sha256:a3459a774d33f5cb63dc451896e065af9661732019057e181a83d5b118f7ebed

Observation ba0676f6-8c18-4488-aa96-6dc25e01e812 · outbound

This paper cites Universal Deep Neural Network Compres- sion.IEEE Journal of Selected Topics in Signal Processing, 14(4):715–726, May 2020.

Reducing Storage of Pretrained Neural Networks by Rate-Constrained Quantization and Entropy Coding Universal Deep Neural Network Compres- sion.IEEE Journal of Selected Topics in Signal Processing, 14(4):715–726, May 2020

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:32:27.628117Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:32:16.835954Z digest=sha256:35defccf0e59831e4ba9cf41aa6c334af4936f3715ca1051faa2b224bd9cc53f

Observation 2c2e90cb-97c4-455d-b258-6aac7cd8bb31 · outbound

This paper cites Imagenet: A large- scale hierarchical image database.

Reducing Storage of Pretrained Neural Networks by Rate-Constrained Quantization and Entropy Coding Imagenet: A large- scale hierarchical image database

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-07T14:32:16.916800Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:32:16.916800Z digest=sha256:4dcadb4d3e364bbf97f217a30e514fb4627e6bb2d5df75c032aa3cee25f5e993

Observation 2c31a623-0c1b-40ab-8873-f52a6e545d96 · outbound

This paper cites GPT3.int8(): 8-bit Matrix Multiplication for Transformers at Scale.Neural Information Processing Systems (NeurIPS), January 2022.

Reducing Storage of Pretrained Neural Networks by Rate-Constrained Quantization and Entropy Coding GPT3.int8(): 8-bit Matrix Multiplication for Transformers at Scale.Neural Information Processing Systems (NeurIPS), January 2022

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:32:27.442812Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:32:16.971706Z digest=sha256:a599099a2be2ffbea379e9c82395591231f2163f42ca65f4d816455576219562

Observation b61ccbcd-d1ed-45a4-96b9-10487a13ebdc · outbound

This paper cites QLoRA: Efficient Finetuning of Quantized LLMs, May 2023.

Reducing Storage of Pretrained Neural Networks by Rate-Constrained Quantization and Entropy Coding QLoRA: Efficient Finetuning of Quantized LLMs, May 2023

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:32:27.226382Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:32:17.038296Z digest=sha256:79fbc2b56273e7541a3897ba72f48bdd808a1f7ab9461c9cc36168531c690cfa

Observation 08b72a21-f583-4eba-b37f-99687fdc40e1 · outbound

This paper cites Svirschevski, Vage Egiazarian, Denis Kuznedelev, Elias Frantar, Saleh Ashkboos, Alexander Borzunov, Torsten Hoefler, and Dan Alistarh.

Reducing Storage of Pretrained Neural Networks by Rate-Constrained Quantization and Entropy Coding Svirschevski, Vage Egiazarian, Denis Kuznedelev, Elias Frantar, Saleh Ashkboos, Alexander Borzunov, Torsten Hoefler, and Dan Alistarh

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:32:27.050348Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:32:17.101095Z digest=sha256:e53e3b7843679ce22ea8006dfbb8606af0c057d4f9a632976d6a5ffc4b9d1b7c

Observation 7cb383c7-385d-49cc-8c08-2913a1704ae8 · outbound

This paper cites The case for 4-bit precision: K-bit Inference Scaling Laws.

Reducing Storage of Pretrained Neural Networks by Rate-Constrained Quantization and Entropy Coding The case for 4-bit precision: K-bit Inference Scaling Laws

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:32:26.869093Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:32:17.167832Z digest=sha256:87b654661a79d365e1701bfe00a3c2eae7b7ac77db39224cbd0fdd0b66b17574

Observation efdb50bb-466a-4a36-b625-99f1d7e75c22 · outbound

This paper cites STBLLM: Breaking the 1-Bit Barrier with Structured Binary LLMs, August 2024.

Reducing Storage of Pretrained Neural Networks by Rate-Constrained Quantization and Entropy Coding STBLLM: Breaking the 1-Bit Barrier with Structured Binary LLMs, August 2024

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:32:26.747986Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:32:17.246176Z digest=sha256:e2e826c6e351a621d0d1f60e488d70b3d0f4fb019a842e8819f264a39eed9ab8

Observation 50f7a082-196c-4b17-b9ab-92101700f1d1 · outbound

This paper cites The use of asymmetric numeral systems as an accurate replacement for huffman coding.

Reducing Storage of Pretrained Neural Networks by Rate-Constrained Quantization and Entropy Coding The use of asymmetric numeral systems as an accurate replacement for huffman coding

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:32:26.630506Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:32:17.325605Z digest=sha256:f9d76fb4c98e9cf21d129828c499753930e896dcbecf8744947b3b3fe666f984

Observation 966c4542-3a45-410b-93a2-27b59bc06d25 · outbound

This paper cites Extreme Compression of Large Language Models via Additive Quantization, September 2024.

Reducing Storage of Pretrained Neural Networks by Rate-Constrained Quantization and Entropy Coding Extreme Compression of Large Language Models via Additive Quantization, September 2024

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:32:26.543235Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:32:17.399995Z digest=sha256:68e5428814f22641f7c3951a083a2244ef8230e73a82181b621711d752cbe7cf

Observation a7835023-4e6f-4a04-b6a4-30af7208a150 · outbound

This paper cites Optimal Brain Compression: A Framework for Accurate Post- Training Quantization and Pruning.

Reducing Storage of Pretrained Neural Networks by Rate-Constrained Quantization and Entropy Coding Optimal Brain Compression: A Framework for Accurate Post- Training Quantization and Pruning

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:32:26.389323Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:32:17.461915Z digest=sha256:275b420ad06b0534c532ea89d8a6ee2a5ed0dcdfb4982ff862c2edaff42ed203

Observation 97d2e906-e906-4fd3-a48c-fb5c6128f318 · outbound

This paper cites SparseGPT: Massive Language Models Can Be Accurately Pruned in One-Shot, March 2023.

Reducing Storage of Pretrained Neural Networks by Rate-Constrained Quantization and Entropy Coding SparseGPT: Massive Language Models Can Be Accurately Pruned in One-Shot, March 2023

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:32:26.279903Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:32:17.553273Z digest=sha256:1c79363996f798d80b9db1b7e863ba4a1cbdb05888ffa8168da672de58d1fbef

Observation 11561f6f-8600-4d3d-bc06-6a2290aac634 · outbound

This paper cites OPTQ: Accurate quan- tization for generative pre-trained transformers.

Reducing Storage of Pretrained Neural Networks by Rate-Constrained Quantization and Entropy Coding OPTQ: Accurate quan- tization for generative pre-trained transformers

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:32:26.200374Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:32:17.633391Z digest=sha256:3ca6bcc302cc50424834f50380f59658531f716a32bbde343b15538db72c948e

Observation a0f1d2a1-98dd-4ac6-8a23-f0827d426d36 · outbound

This paper cites Compression Scaling Laws:Unifying Sparsity and Quantization, February 2025.

Reducing Storage of Pretrained Neural Networks by Rate-Constrained Quantization and Entropy Coding Compression Scaling Laws:Unifying Sparsity and Quantization, February 2025

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:32:26.118695Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:32:17.690251Z digest=sha256:74ffb73952bf7a101db4d3deab33b756e20be7c3ebd2e0416f7ca151c468ed82

Observation 3805198f-2f45-4af3-a9e4-4bd38bd0eb37 · outbound

This paper cites MiniLLM: Knowledge Distillation of Large Language Models.

Reducing Storage of Pretrained Neural Networks by Rate-Constrained Quantization and Entropy Coding MiniLLM: Knowledge Distillation of Large Language Models

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:32:25.979526Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:32:17.769125Z digest=sha256:eff9a80bd715bb58aa8eea3bb8239898d98607458b73f84ed4fd6c1ccad95fe7

Observation e19cf533-d423-479f-9af8-f067ea7af6d8 · outbound

This paper cites an unresolved cited work.

Reducing Storage of Pretrained Neural Networks by Rate-Constrained Quantization and Entropy Coding Unresolved cited work

Reference 23

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:32:25.863479Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:32:17.836993Z digest=sha256:9ffe8b03d0668c158bb12ae21b8286764ca4820d0fe13c3144b24395a5db7dac

Observation 29015da6-6314-4e3e-862d-e23f06486f50 · outbound

This paper cites NeuZip: Memory-Efficient Training and Inference with Dynamic Compression of Neural Networks, October 2024.

Reducing Storage of Pretrained Neural Networks by Rate-Constrained Quantization and Entropy Coding NeuZip: Memory-Efficient Training and Inference with Dynamic Compression of Neural Networks, October 2024

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:32:25.743356Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:32:17.905015Z digest=sha256:40bc0fc517a76d51d1d1eb61867dea78dda488f3337e9b9fb8fca60719bf6ce9

Observation b56910ea-1476-4b58-98fd-403b3e979413 · outbound

This paper cites Hassibi, D.G.

Reducing Storage of Pretrained Neural Networks by Rate-Constrained Quantization and Entropy Coding Hassibi, D.G

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:32:25.630947Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:32:17.992323Z digest=sha256:bacae689d85eac8857d078def83c8c2e6ceb86a3afdc959d1929b34a28097bcc

Observation c06c72ef-c27c-4b0d-9d36-b769f6272758 · outbound

This paper cites Deep Residual Learning for Image Recognition.

Reducing Storage of Pretrained Neural Networks by Rate-Constrained Quantization and Entropy Coding Deep Residual Learning for Image Recognition

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:32:25.469500Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:32:18.053367Z digest=sha256:891ebd86b5d0c502d5aa915897cef95083d5e240cadf71d442a67ff53b428658

Observation f52202bb-ac7e-4232-94c8-eca58327caa6 · outbound

This paper cites Model Compression in Practice: Lessons Learned from Practitioners Creating On-device Machine Learning Experi- ences.

Reducing Storage of Pretrained Neural Networks by Rate-Constrained Quantization and Entropy Coding Model Compression in Practice: Lessons Learned from Practitioners Creating On-device Machine Learning Experi- ences

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:32:25.359504Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:32:18.110251Z digest=sha256:e46ae9c3d490a952aa21fd4a725376342a23b366697268cc457ae42da50707ba

Observation 04780eec-bda4-4ae5-a8b2-4e05c0ba1871 · outbound

This paper cites Mahoney, Yakun Sophia Shao, Kurt Keutzer, and Amir Gholami.

Reducing Storage of Pretrained Neural Networks by Rate-Constrained Quantization and Entropy Coding Mahoney, Yakun Sophia Shao, Kurt Keutzer, and Amir Gholami

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:32:25.189720Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:32:18.187867Z digest=sha256:48fc2834f6fb4cd6a23d4a8f52d363641c47b5e91d62f5b8545deeeb9d6bac82

Observation 07cebcb8-7ea8-4c33-a4ed-fc6ea06072be · outbound

This paper cites Le, and Hartwig Adam.

Reducing Storage of Pretrained Neural Networks by Rate-Constrained Quantization and Entropy Coding Le, and Hartwig Adam

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:32:25.077667Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:32:18.289283Z digest=sha256:fc8587a945c9399ba21ee89cceac6b96c0f84386dc1ee0804c2f95048d9c07ed

Observation 9a27fcbb-35a0-41ce-8a89-2b75d6c075cc · outbound

This paper cites Accurate Post Training Quantization With Small Calibration Sets.

Reducing Storage of Pretrained Neural Networks by Rate-Constrained Quantization and Entropy Coding Accurate Post Training Quantization With Small Calibration Sets

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:32:24.916723Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:32:18.361634Z digest=sha256:5e92d964da6f32f93d950feacb4f41f46fe82941b958946f035ad27371c4ae93

Observation cb6914b4-e72e-44da-912b-88c05c479446 · outbound

This paper cites Mahoney, and Kurt Keutzer.

Reducing Storage of Pretrained Neural Networks by Rate-Constrained Quantization and Entropy Coding Mahoney, and Kurt Keutzer

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:32:24.770296Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:32:18.411007Z digest=sha256:2e9617328836c12d9652442da2c644e90d953bee720fe1aa1c1f2d7831a17cfa

Observation 7b2c153c-0e14-4d83-8602-23851cdba37d · outbound

This paper cites Aksu, Miska M.

Reducing Storage of Pretrained Neural Networks by Rate-Constrained Quantization and Entropy Coding Aksu, Miska M

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:32:24.589078Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:32:18.463769Z digest=sha256:3cefa007c664605d533c25e4baf72d63d079953eb71b8d9224509c95918a483e

Observation 678cbc9e-cf74-4e3f-9650-dce56d55ccca · outbound

This paper cites Adaptive weight compression for memory-efficient neural networks.

Reducing Storage of Pretrained Neural Networks by Rate-Constrained Quantization and Entropy Coding Adaptive weight compression for memory-efficient neural networks

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:32:24.467567Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:32:18.535752Z digest=sha256:1fdcee05d5fb3046e0ce08f96ad4da00228379c374ded7038a92dfc825acc2b7

Observation 0c3747bb-9ab5-48b5-95a3-dcb1e6d96250 · outbound

This paper cites Cifar-10 (canadian institute for advanced research).

Reducing Storage of Pretrained Neural Networks by Rate-Constrained Quantization and Entropy Coding Cifar-10 (canadian institute for advanced research)

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-07T14:32:18.611847Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:32:18.611847Z digest=sha256:5469facc8217609cfca7439a9ed3419dd2e5882a04eb099c5880d9af05d08a55

Observation fedf0c03-7b2e-4efa-b589-abb654016c75 · outbound

This paper cites Energy-Efficient Model Compression and Splitting for Collaborative Inference Over Time-Varying Channels.

Reducing Storage of Pretrained Neural Networks by Rate-Constrained Quantization and Entropy Coding Energy-Efficient Model Compression and Splitting for Collaborative Inference Over Time-Varying Channels

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:32:24.309768Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:32:18.661478Z digest=sha256:5883b35f3df546f0ba3d3b5f19c67899e45675bb5bd5717c0cbec60a9f516b50

Observation f19237d7-95f6-4def-ad28-268ca16b2138 · outbound

This paper cites Gonzalez, Hao Zhang, and Ion Stoica.

Reducing Storage of Pretrained Neural Networks by Rate-Constrained Quantization and Entropy Coding Gonzalez, Hao Zhang, and Ion Stoica

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-07T14:32:18.725195Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:32:18.725195Z digest=sha256:636f3c233af72cbad8e59c122c2ad3545c06749587164973cd5a3031c123b9b2

Observation 49b08133-2996-49c4-9b5a-a3a463600a6f · outbound

This paper cites Memory Efficient Optimizers with 4-bit States.Advances in Neural Information Processing Systems, 36:15136–15171, December 2023.

Reducing Storage of Pretrained Neural Networks by Rate-Constrained Quantization and Entropy Coding Memory Efficient Optimizers with 4-bit States.Advances in Neural Information Processing Systems, 36:15136–15171, December 2023

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:32:24.193718Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:32:18.777724Z digest=sha256:3c6515f7f8b8882b9d7711f422920bd5630b74673567167f52f20d01e6ff8996

Observation ea0ff6d2-256e-4297-bb5e-61e19044e52c · outbound

This paper cites PENNI: Pruned Kernel Sharing for Efficient CNN Inference.

Reducing Storage of Pretrained Neural Networks by Rate-Constrained Quantization and Entropy Coding PENNI: Pruned Kernel Sharing for Efficient CNN Inference

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:32:24.038562Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:32:18.879989Z digest=sha256:934b9aff3253077b10188d392add4170d78cf65a4cd201c8709decd78382f256

Observation acbd46d8-c6fd-4c0b-9d62-f3bfd2794fb1 · outbound

This paper cites AWQ: Activation-aware Weight Quantization for LLM Compression and Acceleration, April 2024.

Reducing Storage of Pretrained Neural Networks by Rate-Constrained Quantization and Entropy Coding AWQ: Activation-aware Weight Quantization for LLM Compression and Acceleration, April 2024

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:32:23.928089Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:32:18.942144Z digest=sha256:2705cd96b7de96079b9815bca9206755a172faec8f641afd545a51b453d7330e

Observation 47fc9a82-5229-4550-931c-928db5094e91 · outbound

This paper cites Cambridge university press, 2003.

Reducing Storage of Pretrained Neural Networks by Rate-Constrained Quantization and Entropy Coding Cambridge university press, 2003

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:32:23.749967Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:32:19.009752Z digest=sha256:eab855406e8a05d7599379ccb5cd79a0b5d21a719a1d532731f768c092c39055

Observation 2812bc00-7f8b-45f5-a989-3b49903e3ae6 · outbound

This paper cites Range encoding: an algorithm for removing redundancy from a digitised message.

Reducing Storage of Pretrained Neural Networks by Rate-Constrained Quantization and Entropy Coding Range encoding: an algorithm for removing redundancy from a digitised message

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:32:23.570244Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:32:19.078862Z digest=sha256:9c1825b1ae6aec9e48b20e0514521fc5fcdc01b8c05778a868d0b8792a2f1684

Observation b6c3659e-7fcb-4b0e-baf2-a7b7ec128f50 · outbound

This paper cites Up or Down? Adaptive Rounding for Post-Training Quantization.

Reducing Storage of Pretrained Neural Networks by Rate-Constrained Quantization and Entropy Coding Up or Down? Adaptive Rounding for Post-Training Quantization

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:32:23.422163Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:32:19.142752Z digest=sha256:7f437180fbba7d0d36047b85d0dd28e0d3ae1e6c4c361acb51fcf4da11741a56

Observation 82cadcdb-dc92-4d6c-a9ed-4b73d51897fe · outbound

This paper cites A White Paper on Neural Network Quantization, June 2021.

Reducing Storage of Pretrained Neural Networks by Rate-Constrained Quantization and Entropy Coding A White Paper on Neural Network Quantization, June 2021

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:32:23.259329Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:32:19.217951Z digest=sha256:8ec06f6f8b5e98f17fb89517552d905b793229cbf2f1f1d61ccabcc78b5f09c9

Observation 83cc05eb-0e6b-4970-9389-f3788a22d706 · outbound

This paper cites Kübler, Jiaji Huang, Matthäus Kleindessner, Jun Huan, V olkan Cevher, Yida Wang, and George Karypis.

Reducing Storage of Pretrained Neural Networks by Rate-Constrained Quantization and Entropy Coding Kübler, Jiaji Huang, Matthäus Kleindessner, Jun Huan, V olkan Cevher, Yida Wang, and George Karypis

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:32:23.089866Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:32:19.307019Z digest=sha256:b57a29488a973b59f756bfe54370375f2940af806e2b86e72abedfeaef85659f

Observation 0a6ff35c-4c2c-4566-8937-aa287d1d91a3 · outbound

This paper cites PhD thesis, Stanford University CA, 1976.

Reducing Storage of Pretrained Neural Networks by Rate-Constrained Quantization and Entropy Coding PhD thesis, Stanford University CA, 1976

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:32:22.904461Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:32:19.342907Z digest=sha256:512dc4ecc19a6d4deacb56150180185a582d2f2c997d2ff9b9ebbfe5d3eed0cc

Observation 1bd47c1e-0df3-47ac-8a80-3fd74b0e4fb3 · outbound

This paper cites PyTorch: An Imperative Style, High-Performance Deep Learning Library, December 2019.

Reducing Storage of Pretrained Neural Networks by Rate-Constrained Quantization and Entropy Coding PyTorch: An Imperative Style, High-Performance Deep Learning Library, December 2019

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:32:22.735567Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:32:19.398920Z digest=sha256:cde66e8df4051a91de245e81f92ede3d81465c6fac8465518acdee6547a79e11

Observation 8f6cb52b-a8e6-426b-84ce-e94e6f2d0d79 · outbound

This paper cites Accurate LoRA-Finetuning Quantization of LLMs via Information Retention, May 2024.

Reducing Storage of Pretrained Neural Networks by Rate-Constrained Quantization and Entropy Coding Accurate LoRA-Finetuning Quantization of LLMs via Information Retention, May 2024

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:32:22.532664Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:32:19.468784Z digest=sha256:e7fbfc21fd7b602721aacac96811b685c597949473b4640b7bb24816b630408f

Observation d191c42e-5b74-4010-8c30-85b192be3aa7 · outbound

This paper cites Arithmetic coding.IBM Journal of research and development, 23(2):149–162, 1979.

Reducing Storage of Pretrained Neural Networks by Rate-Constrained Quantization and Entropy Coding Arithmetic coding.IBM Journal of research and development, 23(2):149–162, 1979

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:32:22.305396Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:32:19.525797Z digest=sha256:55c32e0955d6588cafe329cfb6218422971fadb1a06b67f4bec40f3dfd3796cf

Observation 8e4c1547-574d-463d-8438-b1084a7d111c · outbound

This paper cites an unresolved cited work.

Reducing Storage of Pretrained Neural Networks by Rate-Constrained Quantization and Entropy Coding Unresolved cited work

Reference 49

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:32:22.087711Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:32:19.581936Z digest=sha256:57d0733168c182c93434858d5bfc472ff08a306653a2db1a42d18cf21a5d1ba9

Observation 6a844d03-f3cc-46d2-b71b-58b23dce5c2c · outbound

This paper cites an unresolved cited work.

Reducing Storage of Pretrained Neural Networks by Rate-Constrained Quantization and Entropy Coding Unresolved cited work

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-07T14:32:19.697752Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:32:19.697752Z digest=sha256:332fd1ceb241d3ce425a99b211dd9c4e14dea37b7b4724e5bedc5e57d6300d68

Observation dfbe890b-d91c-4e24-ab30-706c230288ce · outbound

This paper cites FlexGen: High-Throughput Generative Infer- ence of Large Language Models with a Single GPU.

Reducing Storage of Pretrained Neural Networks by Rate-Constrained Quantization and Entropy Coding FlexGen: High-Throughput Generative Infer- ence of Large Language Models with a Single GPU

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:32:21.889948Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:32:19.716104Z digest=sha256:d67485adf5a99fb6114f45119e19a381da2c25168bd1ac18d71167247f85867e

Observation a30e67ad-c023-49aa-b021-828605eae64c · outbound

This paper cites Very Deep Convolutional Networks for Large-Scale Image Recognition.

Reducing Storage of Pretrained Neural Networks by Rate-Constrained Quantization and Entropy Coding Very Deep Convolutional Networks for Large-Scale Image Recognition

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:32:21.699086Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:32:19.719781Z digest=sha256:b259b9eff2440ee0ae0ec82e5ce0efe2a169c94bab154fa368289508016bac9b

Observation 72310331-4608-482e-aaa2-4249c9bdf1dc · outbound

This paper cites Zico Kolter.

Reducing Storage of Pretrained Neural Networks by Rate-Constrained Quantization and Entropy Coding Zico Kolter

Reference 53

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:32:21.525600Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:32:19.756556Z digest=sha256:cf008dc3f6b6219030fcee58065dc50773e2f08bc3daee9beae3e40c0c270a28

Observation 7748fb8f-49b8-4b7e-9827-2715c1895a00 · outbound

This paper cites QuIP#: Even Better LLM Quantization with Hadamard Incoherence and Lattice Codebooks, June 2024.

Reducing Storage of Pretrained Neural Networks by Rate-Constrained Quantization and Entropy Coding QuIP#: Even Better LLM Quantization with Hadamard Incoherence and Lattice Codebooks, June 2024

Reference 54

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:32:21.316087Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:32:19.797300Z digest=sha256:b8d7d8627f49198fc14372282029f50b81656b588fef7c3a96e711f7d8d86625

Observation caf8093d-eec4-4c4a-9e22-05a1ea09391a · outbound

This paper cites DeepCABAC: A Universal Compression Algorithm for Deep Neural Networks.IEEE Journal of Selected Topics in Signal Processing, 14(4):700–714, May 2020.

Reducing Storage of Pretrained Neural Networks by Rate-Constrained Quantization and Entropy Coding DeepCABAC: A Universal Compression Algorithm for Deep Neural Networks.IEEE Journal of Selected Topics in Signal Processing, 14(4):700–714, May 2020

Reference 55

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:32:21.114329Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:32:19.843328Z digest=sha256:3d04e482c65b99b8639520598ccd58498bf34b2cccdd0cf62bc338283866d69f

Observation 11aceea7-f138-4598-bd53-90a71d68a4c1 · outbound

This paper cites Compact and computationally efficient representation of deep neural networks.IEEE Transactions on Neural Networks and Learning Systems, 31(3):772–785, 2020.

Reducing Storage of Pretrained Neural Networks by Rate-Constrained Quantization and Entropy Coding Compact and computationally efficient representation of deep neural networks.IEEE Transactions on Neural Networks and Learning Systems, 31(3):772–785, 2020

Reference 56

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:32:20.902614Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:32:19.886008Z digest=sha256:335e5e37e1014e6af1c2fc930344ac6314083c0cde984b92c5ca178987fd2450

Observation 085bb2b5-222b-4ca7-90dd-5f27afcb3923 · outbound

This paper cites Variational bayesian quantization.

Reducing Storage of Pretrained Neural Networks by Rate-Constrained Quantization and Entropy Coding Variational bayesian quantization

Reference 57

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:32:20.620742Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:32:19.927150Z digest=sha256:f851352279cdba18a265ea4c7277c7d856dc1d08adf7773aab880b17135fb3cc

Observation cc116959-51fc-4bd1-8ffe-bdb6979cacb7 · outbound

This paper cites optimally compensate.

Reducing Storage of Pretrained Neural Networks by Rate-Constrained Quantization and Entropy Coding optimally compensate

Reference 58

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:32:20.341654Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:32:20.014194Z digest=sha256:d357cefd51aea6b2071bf73d4b08bbefe2cfe7393d01f9c0f5146a27ca00574b

Pith citing papers

No inbound Pith citation observations are available.