Pith. sign in

Paper Citation Record · LEDGER

Low-bit Model Quantization for Deep Neural Networks: A Survey

As of 18 August 2026, this Paper Citation Record lists 100 of 115 outbound references and 8 inbound Pith citation observations for arXiv:2505.05530.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.05530 v1

Coverage vector

measured 100 of 115 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-15T23:13:19.861392Z

measured 108 of 108 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-18T06:34:40.430872+00:00

measured 8 of 8 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-07-12T11:46:50.520083Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-01T15:45:48.761436Z

Reference resolution

100 of 115 outbound references displayed

  • verified exact0
  • verified fuzzy46
  • unresolved54
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation d4c433dc-e64d-4bc7-8dc8-955244e2d58d · outbound

This paper cites Pointer sentinel mixture models,.

Low-bit Model Quantization for Deep Neural Networks: A Survey Pointer sentinel mixture models,

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-15T23:13:19.301573Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:13:19.301573Z digest=sha256:c091fd179bd5a086f9b3fe59dc02c4543b6d8ae90bf223800bae4af7364f92da

Observation 7c9f85dd-55e2-4cbc-b044-6ecf5af51e1d · outbound

This paper cites Exploring the limits of transfer learning with a unified text-to-text transformer,.

Low-bit Model Quantization for Deep Neural Networks: A Survey Exploring the limits of transfer learning with a unified text-to-text transformer,

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-15T23:13:19.307627Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:13:19.307627Z digest=sha256:fef3fe4503a4c08325b7f3f90748828c5c49def444dc4be2e0a1d9c0aa66184e

Observation ad231187-c68c-4996-bcd4-223664066619 · outbound

This paper cites Training a Helpful and Harmless Assistant with Reinforcement Learning from Human Feedback.

Low-bit Model Quantization for Deep Neural Networks: A Survey Training a Helpful and Harmless Assistant with Reinforcement Learning from Human Feedback

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-15T23:13:19.313653Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:13:19.313653Z digest=sha256:00b6851ef7d87281d72bafec10d989b2dba0e5ea87edbafadedcd5c87a1c5eef

Observation de0fffe7-7431-4478-bcd4-a6bc94ac23b3 · outbound

This paper cites Stanford alpaca: An instruction- following llama model,.

Low-bit Model Quantization for Deep Neural Networks: A Survey Stanford alpaca: An instruction- following llama model,

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-15T23:13:19.320456Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:13:19.320456Z digest=sha256:9109ed2938d85424fdcb7ce182e577a3ee2bcecf15ae876efa8e81684d04888e

Observation 7c8247db-e210-48d4-bced-65a5883417c7 · outbound

This paper cites The flan collection: Designing data and methods for effective instruction tuning,.

Low-bit Model Quantization for Deep Neural Networks: A Survey The flan collection: Designing data and methods for effective instruction tuning,

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-15T23:13:19.325860Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:13:19.325860Z digest=sha256:febf9b44f0a3ffe07fb2dcdfb0094d147a067a247eb0749ec814719146fb633d

Observation 0b3dec0d-62e1-48f5-aad9-f71b897218ac · outbound

This paper cites Squad: 100,000+ questions for machine comprehension of text,.

Low-bit Model Quantization for Deep Neural Networks: A Survey Squad: 100,000+ questions for machine comprehension of text,

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-15T23:13:19.331391Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:13:19.331391Z digest=sha256:6b54172ba9d31dd82cfffae8755f40d8181428cc12bff6018295b9a5479b959e

Observation ccd6baba-56b0-4b14-845a-8a3a7ccc11ee · outbound

This paper cites The penn treebank: annotating predicate argument structure,.

Low-bit Model Quantization for Deep Neural Networks: A Survey The penn treebank: annotating predicate argument structure,

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-15T23:13:19.337284Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:13:19.337284Z digest=sha256:d1983093dfd618470ca51b234b0ab5c1352334b99e281d9c57476b9d72a05c56

Observation 08cb68cb-d43c-495e-88ff-37b35be21364 · outbound

This paper cites Glue: A multi-task benchmark and analysis platform for natural language understanding,.

Low-bit Model Quantization for Deep Neural Networks: A Survey Glue: A multi-task benchmark and analysis platform for natural language understanding,

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-15T23:13:19.342577Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:13:19.342577Z digest=sha256:98214a3f23f85c903493181825781cf282efdd67ce97a86a5474abf6120eaabf

Observation 6fe00d19-7168-4563-bb3a-87f0cb87066d · outbound

This paper cites Super-naturalinstructions: Generalization via declarative instructions on 1600+ nlp tasks,.

Low-bit Model Quantization for Deep Neural Networks: A Survey Super-naturalinstructions: Generalization via declarative instructions on 1600+ nlp tasks,

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-15T23:13:19.347448Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:13:19.347448Z digest=sha256:fd770889c67e4c7079865c95c4943664038a1165c482f16eb9442fc29e7757ab

Observation 16d7adfe-f2f4-4780-ad5b-3618f7971f4d · outbound

This paper cites Piqa: Reasoning about physical commonsense in natural language,.

Low-bit Model Quantization for Deep Neural Networks: A Survey Piqa: Reasoning about physical commonsense in natural language,

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-15T23:13:19.353409Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:13:19.353409Z digest=sha256:37cfa0d5a36d0207743007603043431b94ab5ee2718201a8f7077949e60a0dd8

Observation 84168dd0-f48a-41ac-858c-aa153cd7b27f · outbound

This paper cites Hellaswag: Can a machine really finish your sentence?.

Low-bit Model Quantization for Deep Neural Networks: A Survey Hellaswag: Can a machine really finish your sentence?

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-15T23:13:19.360380Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:13:19.360380Z digest=sha256:8d6fdf6b660c685fcd5ca08262e718a1adb753f6b349e3d389c2450fcd0aea5e

Observation b0ac00a0-d44f-4331-830a-12e3afaaccbc · outbound

This paper cites Wino- grande: An adversarial winograd schema challenge at scale,.

Low-bit Model Quantization for Deep Neural Networks: A Survey Wino- grande: An adversarial winograd schema challenge at scale,

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-15T23:13:19.366110Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:13:19.366110Z digest=sha256:cb708a14f8e48a1f7091b3c6cb63d276c98b1aa38fe2efef4ec6694ec3aa70f8

Observation 1c37cafe-eab3-4842-adaa-16e302cbae4f · outbound

This paper cites Attention is all you need,.

Low-bit Model Quantization for Deep Neural Networks: A Survey Attention is all you need,

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-15T23:13:19.371754Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:13:19.371754Z digest=sha256:e046a21aeb5a8eff02efe076caae93fb185a6f52b596579271443d372f69ad25

Observation 22d826d9-34f1-4800-83f5-b4cc49affddc · outbound

This paper cites OPT: Open Pre-trained Transformer Language Models.

Low-bit Model Quantization for Deep Neural Networks: A Survey OPT: Open Pre-trained Transformer Language Models

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-15T23:13:19.377457Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:13:19.377457Z digest=sha256:4e89eacf5bf2cb06eebc9e042fbc22f56f4a0691e9c89272e3048aaab94a9e44

Observation dec6f935-5ee6-4d0f-91a8-a93cbdb3e141 · outbound

This paper cites BLOOM: A 176B-Parameter Open-Access Multilingual Language Model.

Low-bit Model Quantization for Deep Neural Networks: A Survey BLOOM: A 176B-Parameter Open-Access Multilingual Language Model

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-15T23:13:19.383064Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:13:19.383064Z digest=sha256:a902bb1338c2b9eca902f8d176dd0f5106fee732aad922cb4a286752a05a751f

Observation ca5d9ce6-7ad0-4b55-9359-7c3af10f133a · outbound

This paper cites Glm-130b: An open bilingual pre-trained model,.

Low-bit Model Quantization for Deep Neural Networks: A Survey Glm-130b: An open bilingual pre-trained model,

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-15T23:13:19.388790Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:13:19.388790Z digest=sha256:f2d79f07009da56743ea12cf9f90788764e7a5e08239c2e85acd09dcf9729393

Observation 6bc31284-9388-4808-95ec-b520ae9a210e · outbound

This paper cites Using DeepSpeed and Megatron to Train Megatron-Turing NLG 530B, A Large-Scale Generative Language Model.

Low-bit Model Quantization for Deep Neural Networks: A Survey Using DeepSpeed and Megatron to Train Megatron-Turing NLG 530B, A Large-Scale Generative Language Model

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-15T23:13:19.393781Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:13:19.393781Z digest=sha256:b89c4d917be30aecabbd99d56162f1240dc675a531fb6b7f9c7b27feb6198efd

Observation 3cdf1d09-e01d-4e0b-a0b9-5677059a34f5 · outbound

This paper cites Llama 2: Open Foundation and Fine-Tuned Chat Models.

Low-bit Model Quantization for Deep Neural Networks: A Survey Llama 2: Open Foundation and Fine-Tuned Chat Models

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-15T23:13:19.399411Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:13:19.399411Z digest=sha256:4c7b70bda10b73b978f86dd585dbdbe0e21ff4c1b8a378ae3a1186db824aad7d

Observation df12727a-21cb-46ab-98a4-79ea8d457b2c · outbound

This paper cites The Falcon Series of Open Language Models.

Low-bit Model Quantization for Deep Neural Networks: A Survey The Falcon Series of Open Language Models

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-15T23:13:19.405130Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:13:19.405130Z digest=sha256:3571df338b9bbffa69ff4eb8fb5991a2c51b4360400f2256cdc74733803ad4ee

Observation 3672858d-5088-413b-885b-4370345dedfd · outbound

This paper cites Mixtral of Experts.

Low-bit Model Quantization for Deep Neural Networks: A Survey Mixtral of Experts

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-15T23:13:19.411197Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:13:19.411197Z digest=sha256:6032a5d730b87402529315f5497ab2f296188c6995747edf191ea0ee08cb1b1e

Observation e6babd0f-a2c7-47de-b7f5-7125fe98c1f3 · outbound

This paper cites Bert: Pre- training of deep bidirectional transformers for language under- standing,.

Low-bit Model Quantization for Deep Neural Networks: A Survey Bert: Pre- training of deep bidirectional transformers for language under- standing,

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-15T23:13:19.416763Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:13:19.416763Z digest=sha256:f59556700f64e5b54271ae1c87e86e0ed3979f077118044b1f445f725577e361

Observation 374d54f2-233d-4f67-ab84-6d7a2fa89782 · outbound

This paper cites RoBERTa: A Robustly Optimized BERT Pretraining Approach.

Low-bit Model Quantization for Deep Neural Networks: A Survey RoBERTa: A Robustly Optimized BERT Pretraining Approach

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-15T23:13:19.424439Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:13:19.424439Z digest=sha256:defeeeb8c9efb499ed8a77f0515d691e4bc934c3fe96fc742c2d92a14caad632

Observation f7ea281b-ef24-4523-bc03-cee510a8d9d2 · outbound

This paper cites Xlnet: Generalized autoregressive pretraining for language understanding,.

Low-bit Model Quantization for Deep Neural Networks: A Survey Xlnet: Generalized autoregressive pretraining for language understanding,

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-15T23:13:19.433508Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:13:19.433508Z digest=sha256:9aa6bb67f109e215979aeba5a9cfb50ecd7f7baa6cd863fdefb5d8179a1933c0

Observation c7376160-75d2-4927-957c-6695b45996b5 · outbound

This paper cites Gpt-j-6b: A 6 billion parameter autoregressive language model,.

Low-bit Model Quantization for Deep Neural Networks: A Survey Gpt-j-6b: A 6 billion parameter autoregressive language model,

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-15T23:13:19.440135Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:13:19.440135Z digest=sha256:7db562289c12d33b43ed5113de774db7e5dfbfbda8429edd8b5615992fd7047d

Observation b83c480b-c9e5-416f-abf0-47e0637d48ff · outbound

This paper cites GPT-NeoX-20B: An Open-Source Autoregressive Language Model.

Low-bit Model Quantization for Deep Neural Networks: A Survey GPT-NeoX-20B: An Open-Source Autoregressive Language Model

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-15T23:13:19.445717Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:13:19.445717Z digest=sha256:9bad995ddf103d34edcd2dc865de23f387759ac190fd96f9ac34478691835028

Observation c3ced501-4c15-4227-adbb-fb26f8be21f3 · outbound

This paper cites Mistral 7B.

Low-bit Model Quantization for Deep Neural Networks: A Survey Mistral 7B

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-15T23:13:19.451214Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:13:19.451214Z digest=sha256:7c322fda122e007fe0b7c7f856e2f399109f9043acccbd5c59b7f6eaba56ffdf

Observation 80ac65f1-8fdf-4fb1-b5c6-2b6008d72bf7 · outbound

This paper cites Starcoder: may the source be with you!.

Low-bit Model Quantization for Deep Neural Networks: A Survey Starcoder: may the source be with you!

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-15T23:13:19.457415Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:13:19.457415Z digest=sha256:7e722e140f9a6450b590e4ce2b21ee819187721e335dce254cea32973ff829e6

Observation a91eb1cf-6e9d-44f0-b042-b32e60b4baac · outbound

This paper cites Gemma: Open Models Based on Gemini Research and Technology.

Low-bit Model Quantization for Deep Neural Networks: A Survey Gemma: Open Models Based on Gemini Research and Technology

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-15T23:13:19.466372Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:13:19.466372Z digest=sha256:ef5621aac0363c55e74179ca28dc0969fa2834c47d425587c5138f8d2e2266e8

Observation 1c272069-a21a-4d17-97ef-8304646565ff · outbound

This paper cites Training language models to follow instructions with human feedback,.

Low-bit Model Quantization for Deep Neural Networks: A Survey Training language models to follow instructions with human feedback,

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-15T23:13:19.473232Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:13:19.473232Z digest=sha256:7f5743989051071f4685eac09aebf41561ae6e3332550963576ae33050358125

Observation d31ea53e-08d8-4148-bf08-06079fe7f90d · outbound

This paper cites Zeroquant: Efficient and affordable post-training quantization for large-scale transformers,.

Low-bit Model Quantization for Deep Neural Networks: A Survey Zeroquant: Efficient and affordable post-training quantization for large-scale transformers,

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-15T23:13:19.478220Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:13:19.478220Z digest=sha256:670677cc7f691cfb1dce0b65100b9a7c52efc426484b3d0da98ef222dbc7972d

Observation 7b8f45e1-2105-4996-9752-0fff3816dc08 · outbound

This paper cites Smoothquant: Accurate and efficient post-training quantization for large language models,.

Low-bit Model Quantization for Deep Neural Networks: A Survey Smoothquant: Accurate and efficient post-training quantization for large language models,

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-15T23:13:19.483884Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:13:19.483884Z digest=sha256:dfe33633e4dbaf3228ac395410bebdd113b43996c50dc82518194afd14df2b70

Observation 7dbec1fb-aa5e-4346-88d9-ad669cbc4ed8 · outbound

This paper cites Squeezellm: Dense-and-sparse quanti- zation,.

Low-bit Model Quantization for Deep Neural Networks: A Survey Squeezellm: Dense-and-sparse quanti- zation,

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-15T23:13:19.490788Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:13:19.490788Z digest=sha256:3c78f9fbdf526c89ab05d695acc465f41bdc969a325ec3b1f50c50437314003e

Observation b4314687-587d-484e-adf0-2c223c6b8d54 · outbound

This paper cites Q-bert: Hessian based ultra low precision quantization of bert,.

Low-bit Model Quantization for Deep Neural Networks: A Survey Q-bert: Hessian based ultra low precision quantization of bert,

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-15T23:13:19.496455Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:13:19.496455Z digest=sha256:20d6458b8fa095240962d588ac5422336627c852a7ee2749e42b76bcd8fe0d1a

Observation d05075f5-2395-49c7-be44-9150fd1098e6 · outbound

This paper cites Qlora: Efficient finetuning of quantized llms,.

Low-bit Model Quantization for Deep Neural Networks: A Survey Qlora: Efficient finetuning of quantized llms,

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-15T23:13:19.502538Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:13:19.502538Z digest=sha256:086970fa744808f225d46bd803f2ece4f398b1abf0b8523bab8e8343b39f2610

Observation 51cd4fc9-28f7-41f3-884f-8b4d8126dc16 · outbound

This paper cites Gptq: Accurate post-training quantization for generative pre-trained transformers,.

Low-bit Model Quantization for Deep Neural Networks: A Survey Gptq: Accurate post-training quantization for generative pre-trained transformers,

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-15T23:13:19.508287Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:13:19.508287Z digest=sha256:8a9c8edc91802e76c8a87d39438dd9499195d73841fb5ce66154e2d1d14967b1

Observation 6787e335-578c-4eeb-a480-69c967f29cc5 · outbound

This paper cites Spqr: A sparse-quantized representation for near-lossless llm weight compression,.

Low-bit Model Quantization for Deep Neural Networks: A Survey Spqr: A sparse-quantized representation for near-lossless llm weight compression,

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-15T23:13:19.513481Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:13:19.513481Z digest=sha256:65d753facf337a922257dd66a0ef52eb41f90fe8b457065f85a574d235969fd8

Observation 7d21d452-8d91-4c81-b5ce-5eca502624e4 · outbound

This paper cites Awq: Activation-aware weight quantization for llm compression and acceleration,.

Low-bit Model Quantization for Deep Neural Networks: A Survey Awq: Activation-aware weight quantization for llm compression and acceleration,

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-15T23:13:19.518391Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:13:19.518391Z digest=sha256:85bfe5e89005c3c1a554a0391d8ddbc0a0adb10bd7b874d6693330213708f38c

Observation 1031e4fc-9521-4848-a71e-0e8d7d1ef6c4 · outbound

This paper cites Training with quantization noise for extreme model compression,.

Low-bit Model Quantization for Deep Neural Networks: A Survey Training with quantization noise for extreme model compression,

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-15T23:13:19.523916Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:13:19.523916Z digest=sha256:914e6748a8ac18b59d4d23025b10df66321521b62c0d967253570c6b9e2a2ffe

Observation 3b955b59-d53a-4e20-88f8-87357c151536 · outbound

This paper cites Obelics: An open web-scale filtered dataset of interleaved image-text documents,.

Low-bit Model Quantization for Deep Neural Networks: A Survey Obelics: An open web-scale filtered dataset of interleaved image-text documents,

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-15T23:13:19.528929Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:13:19.528929Z digest=sha256:a3cb3461e2bbfd8f950b15360e3b3a7e8f69e50dae9cc7465047f67cb57f8449

Observation 428ce152-9e8a-45f2-acca-d8a11948464f · outbound

This paper cites Laion-5b: An open large-scale dataset for training next generation image-text models,.

Low-bit Model Quantization for Deep Neural Networks: A Survey Laion-5b: An open large-scale dataset for training next generation image-text models,

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-15T23:13:19.534387Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:13:19.534387Z digest=sha256:fc95a29f1f7ba5d7902d9775afc1babe96b08a06f6545c3522d9485e20bfbe8a

Observation 93f2e3d4-016a-4efd-99db-23e35796e85a · outbound

This paper cites Wit: Wikipedia-based image text dataset for multimodal multilin- gual machine learning,.

Low-bit Model Quantization for Deep Neural Networks: A Survey Wit: Wikipedia-based image text dataset for multimodal multilin- gual machine learning,

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-15T23:13:19.539647Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:13:19.539647Z digest=sha256:ef0bccc27d3503c3e99c079a890622afdb966d39de8185890c6f3b0275ac4688

Observation 2fb15a19-b065-401b-980b-ad80e07494cc · outbound

This paper cites Gqa: A new dataset for real- world visual reasoning and compositional question answering,.

Low-bit Model Quantization for Deep Neural Networks: A Survey Gqa: A new dataset for real- world visual reasoning and compositional question answering,

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-15T23:13:19.544978Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:13:19.544978Z digest=sha256:0478cdc38f6d191303523dc63c4f624cc958509b470f1790226b606725b421b1

Observation 4b9cdd58-e13b-44f4-8fba-c1394432df23 · outbound

This paper cites Towards vqa models that can read,.

Low-bit Model Quantization for Deep Neural Networks: A Survey Towards vqa models that can read,

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-15T23:13:19.550543Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:13:19.550543Z digest=sha256:d7a1a84e3089c610152cef50ec3318ef968aa8e34bd0e842769945fad6fa6651

Observation a33d66cb-8069-4f73-ad12-c5bd4fa41f11 · outbound

This paper cites Learn to explain: Multimodal reasoning via thought chains for science question answering,.

Low-bit Model Quantization for Deep Neural Networks: A Survey Learn to explain: Multimodal reasoning via thought chains for science question answering,

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-15T23:13:19.556053Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:13:19.556053Z digest=sha256:c11285e8a5b2b5d96767088d5c119c964c91b178fef79161c08077e1136c53bc

Observation 64afc370-ae94-47b2-9e70-598a6fd7f437 · outbound

This paper cites Vizwiz grand challenge: Answering visual questions from blind people,.

Low-bit Model Quantization for Deep Neural Networks: A Survey Vizwiz grand challenge: Answering visual questions from blind people,

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-15T23:13:19.561966Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:13:19.561966Z digest=sha256:288149d54df211b48e87ac3a89e00e6f76a1a1da5d4d588a046cccd70c8922db

Observation bab56926-1c31-4e04-8786-205c4bf1c8b2 · outbound

This paper cites Visual instruction tuning,.

Low-bit Model Quantization for Deep Neural Networks: A Survey Visual instruction tuning,

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-15T23:13:19.567377Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:13:19.567377Z digest=sha256:ba2e5df58ae274ebc9b641af08ff3266851e122435ac43a8a2940e0b8233aada

Observation a77461eb-31cd-413c-8188-d06065b286e4 · outbound

This paper cites Openflamingo: An open-source framework for training large autoregressive vision-language models,.

Low-bit Model Quantization for Deep Neural Networks: A Survey Openflamingo: An open-source framework for training large autoregressive vision-language models,

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T23:13:23.567451Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T23:13:19.573035Z digest=sha256:9e000c3d5cc36e5b9c66c8bf2794a569fc33b0f827c5494bef26ea94a5833ec1

Observation 00f8d0aa-8305-4f75-b696-2a5e484d3cb1 · outbound

This paper cites Noisyquant: Noisy bias-enhanced post-training activation quanti- zation for vision transformers,.

Low-bit Model Quantization for Deep Neural Networks: A Survey Noisyquant: Noisy bias-enhanced post-training activation quanti- zation for vision transformers,

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T23:13:23.548794Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T23:13:19.578203Z digest=sha256:f2b5654655c21736af3d8dd8a442156396bd0489afc93aaf41fb35f8b946cb95

Observation 58aec02d-4de8-4902-aa20-68417328c341 · outbound

This paper cites Imagenet large scale visual recognition challenge,.

Low-bit Model Quantization for Deep Neural Networks: A Survey Imagenet large scale visual recognition challenge,

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-15T23:13:19.583664Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:13:19.583664Z digest=sha256:1bfd9cd630931de0a7715e516ebe200491ef8902aab606523e06db603333fe67

Observation 2ef7bc6f-ba16-4f75-b036-44c9495225f9 · outbound

This paper cites Learning multiple layers of features from tiny images,.

Low-bit Model Quantization for Deep Neural Networks: A Survey Learning multiple layers of features from tiny images,

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T23:13:23.520121Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T23:13:19.589121Z digest=sha256:379416306fd2455cb3c613cc9d0407958dd3f52c10741393150d148f2871fad1

Observation 3986a7fd-ed68-4074-8fe4-e42c7337376b · outbound

This paper cites Microsoft COCO Captions: Data Collection and Evaluation Server.

Low-bit Model Quantization for Deep Neural Networks: A Survey Microsoft COCO Captions: Data Collection and Evaluation Server

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-15T23:13:19.594472Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:13:19.594472Z digest=sha256:3b7d9133fb6e8c8e3df94ab803c816d1d387c51b51f30cd986752b98aade7a08

Observation fbda1bb2-0b2b-4d3a-a907-0bf65c6cb46a · outbound

This paper cites LSUN: Construction of a Large-scale Image Dataset using Deep Learning with Humans in the Loop.

Low-bit Model Quantization for Deep Neural Networks: A Survey LSUN: Construction of a Large-scale Image Dataset using Deep Learning with Humans in the Loop

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-15T23:13:19.600703Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:13:19.600703Z digest=sha256:2c75ba584ee21d3a50ee58434674ab7a4f52aab3c493fcce490479890fab8007

Observation b46f1a0a-f61d-4d6b-8576-5654bbf11a35 · outbound

This paper cites Mak- ing the v in vqa matter: Elevating the role of image understanding in visual question answering,.

Low-bit Model Quantization for Deep Neural Networks: A Survey Mak- ing the v in vqa matter: Elevating the role of image understanding in visual question answering,

Reference 53

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T23:13:23.502605Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T23:13:19.606142Z digest=sha256:6e2bc65724269c35c14c8f46c744aca4040dfc05b380278304e9ee15c4a189ef

Observation f8631432-f4db-43a2-af19-e3f25568717c · outbound

This paper cites Improved denoising diffusion probabilistic models,.

Low-bit Model Quantization for Deep Neural Networks: A Survey Improved denoising diffusion probabilistic models,

Reference 54

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T23:13:23.483718Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T23:13:19.611276Z digest=sha256:86211c3e859eff8d37d39f064db0fb8994647b066844dc203d6c78edc137637b

Observation 910598c7-bf96-4d8d-a1bd-d43e646b7538 · outbound

This paper cites Denoising diffusion probabilistic models,.

Low-bit Model Quantization for Deep Neural Networks: A Survey Denoising diffusion probabilistic models,

Reference 55

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T23:13:23.338418Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T23:13:19.616513Z digest=sha256:c7011d8b97e3c5add435967028a934326fd1263ecb8e4cbcc5e50c93db58a7b0

Observation 42bb0e91-f8f9-46f6-b615-ad05928005c6 · outbound

This paper cites Deep residual learning for image recognition,.

Low-bit Model Quantization for Deep Neural Networks: A Survey Deep residual learning for image recognition,

Reference 56

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T23:13:23.264299Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T23:13:19.621356Z digest=sha256:eab7968a154c514d228f513214227b52ffd27dd0d6a7040437a7a0ce1f3b6ce9

Observation e11a9ad8-cc00-43ec-9634-ba36bfa1dd14 · outbound

This paper cites Mobilenetv2: Inverted residuals and linear bottlenecks,.

Low-bit Model Quantization for Deep Neural Networks: A Survey Mobilenetv2: Inverted residuals and linear bottlenecks,

Reference 57

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T23:13:23.248075Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T23:13:19.626410Z digest=sha256:29b48ca1359d2a1589466809819ed43da1bc577ebe0c4458996d845bcf0fdc13

Observation 56d79bee-8932-4958-8873-814a60407cda · outbound

This paper cites A style-based generator archi- tecture for generative adversarial networks,.

Low-bit Model Quantization for Deep Neural Networks: A Survey A style-based generator archi- tecture for generative adversarial networks,

Reference 58

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T23:13:23.202636Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T23:13:19.631601Z digest=sha256:531cd96d1487660d644718a2bb8fb11c627cf51b268622479f4a0dec24bd9cb6

Observation 66fc02ce-a191-477a-aaa3-5be1241f88cc · outbound

This paper cites Vila: On pre-training for visual language models,.

Low-bit Model Quantization for Deep Neural Networks: A Survey Vila: On pre-training for visual language models,

Reference 59

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T23:13:23.061425Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T23:13:19.637654Z digest=sha256:ccb1f1444be1e4ce3e8ffe834e0604b615ca5314586ef0f986ba54bc4ac02360

Observation f7993d47-e9ce-491a-80c1-2e6881ecc1ff · outbound

This paper cites Swin transformer: Hierarchical vision transformer using shifted windows,.

Low-bit Model Quantization for Deep Neural Networks: A Survey Swin transformer: Hierarchical vision transformer using shifted windows,

Reference 60

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T23:13:22.944959Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T23:13:19.642937Z digest=sha256:faa1d5aa55eaef12ea15d42fd26b88c49bf43396003e93e7009aca2b71637293

Observation 0e64118f-248d-4da1-ab8d-8ce10d91b2f3 · outbound

This paper cites Qdrop: Randomly dropping quantization for extremely low-bit post-training quanti- zation,.

Low-bit Model Quantization for Deep Neural Networks: A Survey Qdrop: Randomly dropping quantization for extremely low-bit post-training quanti- zation,

Reference 61

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T23:13:22.926953Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T23:13:19.647856Z digest=sha256:a5647af2e54f8854a15e6424b1aa2081b036f749d7d8e5d74ed159c79e6b55d6

Observation 74fa6a8f-5602-411c-b988-e89c7ab4d4c9 · outbound

This paper cites Post-training quantization on diffusion models,.

Low-bit Model Quantization for Deep Neural Networks: A Survey Post-training quantization on diffusion models,

Reference 62

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T23:13:22.707014Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T23:13:19.653293Z digest=sha256:94e298fb7e906b35e64eecd0b9cb4e6b0d5e4e6a20b7ae248bbe79803f1ca811

Observation d0493f47-99e2-4eaa-bd99-dcc2f06d67b9 · outbound

This paper cites Q-DM: An efficient low-bit quantized diffusion model,.

Low-bit Model Quantization for Deep Neural Networks: A Survey Q-DM: An efficient low-bit quantized diffusion model,

Reference 63

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T23:13:22.622256Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T23:13:19.658909Z digest=sha256:a691de68690f0d121495b02fb6fae8f7ed47911ec66fe76c364ca341b389df9b

Observation 7df7084f-f2a6-4513-a394-5cb90605ecc7 · outbound

This paper cites Accurate post training quantization with small calibration sets,.

Low-bit Model Quantization for Deep Neural Networks: A Survey Accurate post training quantization with small calibration sets,

Reference 64

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T23:13:22.604050Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T23:13:19.665717Z digest=sha256:f1ccda100552a2dcc3bb98631362db828fa4889fffabbb002157e7b35407d1f1

Observation d94c755a-7381-40a5-bd80-22e19f16b029 · outbound

This paper cites Ntire 2017 challenge on single image super-resolution: Dataset and study,.

Low-bit Model Quantization for Deep Neural Networks: A Survey Ntire 2017 challenge on single image super-resolution: Dataset and study,

Reference 65

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T23:13:22.587317Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T23:13:19.672421Z digest=sha256:c4d36cf8ba2cbb5110d8aa96c0d7f6e37e16ee7bb5abef36bdb15a1c446342e7

Observation b9c14db4-e7fe-4816-984b-13b71a5ab5dc · outbound

This paper cites Component divide-and-conquer for real-world image super- resolution,.

Low-bit Model Quantization for Deep Neural Networks: A Survey Component divide-and-conquer for real-world image super- resolution,

Reference 66

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T23:13:22.569533Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T23:13:19.677928Z digest=sha256:36c00beae52ec47aed59d889eb8324a640c82ca993bfd10d465bf541c8f8051e

Observation ccf9d8d6-eef0-48d0-baa5-667f1af9ae2c · outbound

This paper cites Low-complexity single-image super-resolution based on nonnegative neighbor embedding,.

Low-bit Model Quantization for Deep Neural Networks: A Survey Low-complexity single-image super-resolution based on nonnegative neighbor embedding,

Reference 67

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T23:13:22.550019Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T23:13:19.684017Z digest=sha256:b655f836f03c1d645b4ae50584960af8b51a2d7d742bfae056f258d657939512

Observation 1619a4ba-1ee2-4f37-b610-dc2d798df6b4 · outbound

This paper cites On single image scale-up using sparse-representations,.

Low-bit Model Quantization for Deep Neural Networks: A Survey On single image scale-up using sparse-representations,

Reference 68

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T23:13:22.416809Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T23:13:19.689568Z digest=sha256:044de15cf228d6e64dc1674d7b06cc2a4e45863528f252958fcad189ed5162a6

Observation c4423193-ebfc-4b23-ae3d-97a079b34c09 · outbound

This paper cites A database of human segmented natural images and its application to evaluating segmentation algorithms and measuring ecological statistics,.

Low-bit Model Quantization for Deep Neural Networks: A Survey A database of human segmented natural images and its application to evaluating segmentation algorithms and measuring ecological statistics,

Reference 69

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T23:13:22.369441Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T23:13:19.695011Z digest=sha256:22032a5547d8182d1e0d01febaf6f07b3d53a5a50b63090e072f332e0d4833ea

Observation 18b30f33-3014-4a81-bb81-72d82a44be7c · outbound

This paper cites Single image super- resolution from transformed self-exemplars,.

Low-bit Model Quantization for Deep Neural Networks: A Survey Single image super- resolution from transformed self-exemplars,

Reference 70

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T23:13:22.352026Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T23:13:19.700262Z digest=sha256:2adc29976a44b7460be4b043d21684273024f707a0950ef3e324a7f5afa2d308

Observation e0e19478-2439-4718-b04d-a1ebdfc0fe8a · outbound

This paper cites Accurate image super-resolution using very deep convolutional networks,.

Low-bit Model Quantization for Deep Neural Networks: A Survey Accurate image super-resolution using very deep convolutional networks,

Reference 71

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T23:13:22.333916Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T23:13:19.706238Z digest=sha256:f23cf85a0a580c898f883cd092988066c3651274e2392b8005d78a3645dd390a

Observation d7b25a41-7a76-40fc-a669-b1565a4b1681 · outbound

This paper cites Enhanced deep residual networks for single image super-resolution,.

Low-bit Model Quantization for Deep Neural Networks: A Survey Enhanced deep residual networks for single image super-resolution,

Reference 72

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T23:13:22.316244Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T23:13:19.713178Z digest=sha256:26543de3195688263f6672163b5cbf17a0f4cda3aae17699b06e9d5b1838cadf

Observation a7cd9689-07a5-43ae-9dce-1e3a4d9becf5 · outbound

This paper cites Residual dense network for image super-resolution,.

Low-bit Model Quantization for Deep Neural Networks: A Survey Residual dense network for image super-resolution,

Reference 73

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T23:13:22.153553Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T23:13:19.721060Z digest=sha256:9b1f2ec8a7b92af055987ffdbc9015a6c99fbff210605f41a36390c5cc50516e

Observation 402371c4-fecf-4a7c-8fea-3983c8b2dabf · outbound

This paper cites Photo- realistic single image super-resolution using a generative adver- sarial network,.

Low-bit Model Quantization for Deep Neural Networks: A Survey Photo- realistic single image super-resolution using a generative adver- sarial network,

Reference 74

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T23:13:22.087919Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T23:13:19.727526Z digest=sha256:96a480042a21ae5b3052369df9dd8b5cd49998aea936b1a64aa7e8966fecfc4b

Observation fe6c70d7-ff75-42d6-80a9-ae675da94a33 · outbound

This paper cites Daq: Channel-wise distribution-aware quantization for deep image super-resolution networks,.

Low-bit Model Quantization for Deep Neural Networks: A Survey Daq: Channel-wise distribution-aware quantization for deep image super-resolution networks,

Reference 75

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T23:13:22.067706Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T23:13:19.733221Z digest=sha256:b1a7c9f1b60f76033f02769e893f03a7dc7994ce505dc55722c3d314f4dc9a8f

Observation 629789ed-85c1-4c6c-a977-a9d7c7879f90 · outbound

This paper cites Searching for low-bit weights in quantized neural networks,.

Low-bit Model Quantization for Deep Neural Networks: A Survey Searching for low-bit weights in quantized neural networks,

Reference 76

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T23:13:21.885684Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T23:13:19.739106Z digest=sha256:5eb7720cef87dc169c94104e0967799a739442bc2247a16949aaedd7568c4287

Observation 0523555f-7e8d-4629-b374-b02e07ef61ad · outbound

This paper cites Learnable lookup table for neural network quantization,.

Low-bit Model Quantization for Deep Neural Networks: A Survey Learnable lookup table for neural network quantization,

Reference 77

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T23:13:21.791408Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T23:13:19.744293Z digest=sha256:b8215fa65b94dfc53117ea65f5a607e3be325860a84c42edc26f244ab6af4e4c

Observation e31be9f7-2585-42b4-b7dd-b7f254dd9eda · outbound

This paper cites Encoder-decoder with atrous separable convolution for semantic image segmentation,.

Low-bit Model Quantization for Deep Neural Networks: A Survey Encoder-decoder with atrous separable convolution for semantic image segmentation,

Reference 78

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T23:13:21.773345Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T23:13:19.749003Z digest=sha256:b1745a2d6b75a74d3596aa5cd092b029d3c74131111ea6b417a453dac64723d5

Observation 913aba1e-6f9a-4615-8df7-43a5cc94e500 · outbound

This paper cites Repvgg: Making vgg-style convnets great again,.

Low-bit Model Quantization for Deep Neural Networks: A Survey Repvgg: Making vgg-style convnets great again,

Reference 79

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T23:13:21.627957Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T23:13:19.754090Z digest=sha256:acf4699615403cfa3791bb7ac8fc12ad03c86e6dca94633590e936c3fdf92918

Observation 8b37d7d7-1dc6-400f-bf23-5566dca156a8 · outbound

This paper cites Up or down? adaptive rounding for post-training quantization,.

Low-bit Model Quantization for Deep Neural Networks: A Survey Up or down? adaptive rounding for post-training quantization,

Reference 80

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T23:13:21.563277Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T23:13:19.758749Z digest=sha256:c025c1a3c30fb124a890d28524db87c38a901cf4dfe105b14cdb8317d7d60a82

Observation bd612af2-c7b9-4bca-a937-8d14eab618a5 · outbound

This paper cites Designing network design spaces,.

Low-bit Model Quantization for Deep Neural Networks: A Survey Designing network design spaces,

Reference 81

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T23:13:21.545213Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T23:13:19.763858Z digest=sha256:69c383ac41c5a979b52a0609c0b8ff23d0aa4914c01fad5f8ce85f131274577a

Observation 7672931f-7f13-4a5b-be7b-0443310316ce · outbound

This paper cites Mnasnet: Platform-aware neural architecture search for mobile,.

Low-bit Model Quantization for Deep Neural Networks: A Survey Mnasnet: Platform-aware neural architecture search for mobile,

Reference 82

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T23:13:21.526562Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T23:13:19.768886Z digest=sha256:8b857fb074a877b764f65547294a7fa3dae9ed22067e23d7fe9d68876ed04f6d

Observation 3af671cc-f240-4d24-964b-c40529025a1a · outbound

This paper cites Very deep convolutional net- works for large-scale image recognition,.

Low-bit Model Quantization for Deep Neural Networks: A Survey Very deep convolutional net- works for large-scale image recognition,

Reference 83

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T23:13:21.471076Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T23:13:19.773885Z digest=sha256:eb970d53f9d044cc99e39a8a81168266b9c3f25dc5fc4faf33cda3c1fd227e43

Observation 1cfb2fda-c509-463c-bb82-ffc6deb89743 · outbound

This paper cites Rethinking the inception architecture for computer vision,.

Low-bit Model Quantization for Deep Neural Networks: A Survey Rethinking the inception architecture for computer vision,

Reference 84

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T23:13:21.304035Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T23:13:19.778549Z digest=sha256:37d2d2a4e487f88e044a67b742df4672dd87629bd469c7b153341a6d69ddfcea

Observation 4b439f53-6ea3-42f0-a4c9-953faada2203 · outbound

This paper cites Brecq: Pushing the limit of post-training quantization by block reconstruction,.

Low-bit Model Quantization for Deep Neural Networks: A Survey Brecq: Pushing the limit of post-training quantization by block reconstruction,

Reference 85

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T23:13:21.217667Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T23:13:19.783165Z digest=sha256:960974c5a1590c8f97e5b59fff2613fad4416df033800dcba2ca9ebe1d8976f8

Observation 69755a2b-71e4-4b3d-925c-e796e3fd9c3d · outbound

This paper cites Hawq: Hessian aware quantization of neural networks with mixed-precision,.

Low-bit Model Quantization for Deep Neural Networks: A Survey Hawq: Hessian aware quantization of neural networks with mixed-precision,

Reference 86

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T23:13:21.199956Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T23:13:19.788993Z digest=sha256:19cc9d62f6cee2e5c07a5edb09ecc2d85ae73a5b75822ed400a41dc0908f2353

Observation 5c0b2ae5-827d-4b3c-a3dd-8b1dd08c0655 · outbound

This paper cites Overcoming oscillations in quantization-aware training,.

Low-bit Model Quantization for Deep Neural Networks: A Survey Overcoming oscillations in quantization-aware training,

Reference 87

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T23:13:21.104647Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T23:13:19.794294Z digest=sha256:a7989034808ebaf0fe9a0990cbcaa41ee8c88ce6b3ad7c25f865eb7effd22413

Observation 2f1c154f-a8ba-4759-96ad-0668e6a21009 · outbound

This paper cites Focal loss for dense object detection,.

Low-bit Model Quantization for Deep Neural Networks: A Survey Focal loss for dense object detection,

Reference 88

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T23:13:21.023172Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T23:13:19.800613Z digest=sha256:b54855a3e42f4ba9cb935813be180d422c8958128effc15980d7537fcef0fe4b

Observation c69edea0-5bfa-4fec-bba0-34479dc12a9b · outbound

This paper cites ZeroQuant-V2: Exploring Post-training Quantization in LLMs from Comprehensive Study to Low Rank Compensation.

Low-bit Model Quantization for Deep Neural Networks: A Survey ZeroQuant-V2: Exploring Post-training Quantization in LLMs from Comprehensive Study to Low Rank Compensation

Reference 89

Resolution
unresolved
no resolver link, observed 2026-08-15T23:13:19.805481Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:13:19.805481Z digest=sha256:5d8e78ffc0559e49fbaa4abad0f3b81252edc0e98a843da1ff3c422ef28e0be0

Observation ffa7fade-73c5-4f78-87ea-17bb2a875211 · outbound

This paper cites A comprehen- sive survey on model quantization for deep neural networks in image classification,.

Low-bit Model Quantization for Deep Neural Networks: A Survey A comprehen- sive survey on model quantization for deep neural networks in image classification,

Reference 90

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T23:13:21.003205Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T23:13:19.810904Z digest=sha256:9ebac6d8e3fd39b79c732c5f620b01cacd7b6f5c1d076b7ad4849359af538f28

Observation 038e160a-abbb-4a31-bb8d-9e6043671bf5 · outbound

This paper cites A survey of quantization methods for efficient neural network inference,.

Low-bit Model Quantization for Deep Neural Networks: A Survey A survey of quantization methods for efficient neural network inference,

Reference 91

Resolution
unresolved
no resolver link, observed 2026-08-15T23:13:19.815967Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:13:19.815967Z digest=sha256:1f2a75a85d089c2a7fbae6d77bfdf9de7c183d6535b5b5ca48990e2d04de3e8b

Observation 7a4e6d2a-9e2a-46d3-a366-7b37cf5ff968 · outbound

This paper cites Evaluating Quantized Large Language Models.

Low-bit Model Quantization for Deep Neural Networks: A Survey Evaluating Quantized Large Language Models

Reference 92

Resolution
unresolved
no resolver link, observed 2026-08-15T23:13:19.820832Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:13:19.820832Z digest=sha256:28941ad636d2079c969d1890d2f9c5bcf5b8e5903dfe7441d03c2eb7d078e31c

Observation 23c0c9f8-5cc3-43ae-a1ca-0f1bc939f365 · outbound

This paper cites Exploiting LLM Quantization.

Low-bit Model Quantization for Deep Neural Networks: A Survey Exploiting LLM Quantization

Reference 93

Resolution
unresolved
no resolver link, observed 2026-08-15T23:13:19.825641Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:13:19.825641Z digest=sha256:5f0b67582bad98e6ab6a7d6621c332fe668a37d2f82b2f046314b51816f3dd27

Observation 7b8721ce-bbf3-4a53-b4e0-5fcbc4afa193 · outbound

This paper cites Exploring post-training quantization in llms from comprehensive study to low rank compensation,.

Low-bit Model Quantization for Deep Neural Networks: A Survey Exploring post-training quantization in llms from comprehensive study to low rank compensation,

Reference 94

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T23:13:20.969696Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T23:13:19.830284Z digest=sha256:5957f68fdbd23ee58a7570602dbb58ab27c20ca38af08ebd60d4d78c5f676fd9

Observation 399b7d90-901d-4dbf-b466-da2133c0e6a7 · outbound

This paper cites Exploring quantization techniques for large-scale language models: Methods, challenges and future directions,.

Low-bit Model Quantization for Deep Neural Networks: A Survey Exploring quantization techniques for large-scale language models: Methods, challenges and future directions,

Reference 95

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T23:13:20.854574Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T23:13:19.834912Z digest=sha256:672022e72b1a46f50da82b32909d6130867bb8bb5a2b30c46789cfbc3937fe65

Observation 6c765d08-8b49-4e55-982c-f2b39e1461f2 · outbound

This paper cites The case for 4-bit precision: k-bit inference scaling laws,.

Low-bit Model Quantization for Deep Neural Networks: A Survey The case for 4-bit precision: k-bit inference scaling laws,

Reference 96

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T23:13:20.741994Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T23:13:19.839715Z digest=sha256:66b9b88ffe1c33b229b457c25777216d30021feae229831003a231ce80e555d3

Observation d15ba077-7584-4b14-86c8-05739f4400ce · outbound

This paper cites Only train once: A one-shot neural network training and pruning framework,.

Low-bit Model Quantization for Deep Neural Networks: A Survey Only train once: A one-shot neural network training and pruning framework,

Reference 97

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T23:13:20.723893Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T23:13:19.845366Z digest=sha256:4e184b90dcdcdfbcb209a5d27bda37dd923f152123d93d09875661e73e42ce3b

Observation c1b7694b-31bc-4787-8757-b41a6da224e3 · outbound

This paper cites Deephoyer: Learning sparser neural network with differentiable scale-invariant sparsity measures,.

Low-bit Model Quantization for Deep Neural Networks: A Survey Deephoyer: Learning sparser neural network with differentiable scale-invariant sparsity measures,

Reference 98

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T23:13:20.706752Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T23:13:19.850130Z digest=sha256:230f717fdc8a2215a50a3acfece94467a5be828a9c482bb9263e4afbcd2f551f

Observation e3a67481-d2d0-46b9-b5c0-a751206324f3 · outbound

This paper cites Towards Compact ConvNets via Structure-Sparsity Regularized Filter Pruning.

Low-bit Model Quantization for Deep Neural Networks: A Survey Towards Compact ConvNets via Structure-Sparsity Regularized Filter Pruning

Reference 99

Resolution
unresolved
no resolver link, observed 2026-08-15T23:13:19.855265Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:13:19.855265Z digest=sha256:291b25c868b55cd2cd454dd3c67fbd4627370e04f0c9acddd3c2a9e47a7dfebb

Observation 1a68048a-de30-41d3-b058-20a85153a1d9 · outbound

This paper cites Accelerate cnn via recursive bayesian pruning,.

Low-bit Model Quantization for Deep Neural Networks: A Survey Accelerate cnn via recursive bayesian pruning,

Reference 100

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T23:13:20.611473Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T23:13:19.861392Z digest=sha256:c08a6c874c9ce5eb3173344eada2fe9493466da0e8433f446ea5e3519053000c

Pith citing papers

Observation a2ed310a-1707-4cc8-b7c9-254c3c8be47d · inbound

Zero-Shot Quantization via Weight-Space Arithmetic cites this paper.

Zero-Shot Quantization via Weight-Space Arithmetic Low-bit Model Quantization for Deep Neural Networks: A Survey

Reference 4

Resolution
verified exact
arxiv_id, observed 2026-05-13T20:08:12.603473Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-13T20:07:41.196837Z digest=sha256:18302f2f563d90715d81d67ef56b3a3ac8930904a42e3eb42297b20cfb906609

Observation 29cfaa74-df41-4ec8-a78d-fe5a7e64354a · inbound

TinyNeRV: Compact Neural Video Representations via Capacity Scaling, Distillation, and Low-Precision Inference cites this paper.

TinyNeRV: Compact Neural Video Representations via Capacity Scaling, Distillation, and Low-Precision Inference Low-bit Model Quantization for Deep Neural Networks: A Survey

Reference 60

Resolution
verified exact
arxiv_id, observed 2026-05-11T00:30:52.255853Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-10T18:29:43.498730Z digest=sha256:f0f7256b83f6b2ff618eb9e955a8ce3fa2df8e51ea84a494eb54e1cec825079f

Observation d8d30dce-ae4f-479e-a965-2c7e9174fa02 · inbound

DECO: Sparse Mixture-of-Experts with Dense-Comparable Performance on End-Side Devices cites this paper.

DECO: Sparse Mixture-of-Experts with Dense-Comparable Performance on End-Side Devices Low-bit Model Quantization for Deep Neural Networks: A Survey

Reference 179

Resolution
verified exact
arxiv_id, observed 2026-05-12T03:36:20.006448Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-05-12T03:36:12.915133Z digest=sha256:108f831b6df23915537a7e5320249123de5d9668b781c8c86644efb735eb952f

Observation 01dbdfb9-1caf-4e20-9e3d-ba1aa9cd5d94 · inbound

DECO: Sparse Mixture-of-Experts with Dense-Comparable Performance on End-Side Devices cites this paper.

DECO: Sparse Mixture-of-Experts with Dense-Comparable Performance on End-Side Devices Low-bit Model Quantization for Deep Neural Networks: A Survey

Reference 179

Resolution
verified exact
arxiv_id, observed 2026-05-13T07:32:30.232726Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-05-13T07:29:14.545746Z digest=sha256:e880ade69b2de9d02a59ae74a7962d286879a413102291ec3f21c5a8baa13e2c

Observation 088da7db-f95a-4ad4-9896-94c0b86e197a · inbound

DECO: Sparse Mixture-of-Experts with Dense-Comparable Performance on End-Side Devices cites this paper.

DECO: Sparse Mixture-of-Experts with Dense-Comparable Performance on End-Side Devices Low-bit Model Quantization for Deep Neural Networks: A Survey

Reference 179

Resolution
verified exact
arxiv_id, observed 2026-05-21T07:59:50.126006Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-05-21T07:57:49.746594Z digest=sha256:8366959e7e9d6155fb8cedba576c41de5a7c2bf9aec2cde93c38094e841646b9

Observation d241124c-6c77-482f-a11a-e77cc63a4737 · inbound

MARR: Module-Adaptive Residual Reconstruction for Low-Bit Post-Training Quantization cites this paper.

MARR: Module-Adaptive Residual Reconstruction for Low-Bit Post-Training Quantization Low-bit Model Quantization for Deep Neural Networks: A Survey

Reference 22

Resolution
verified exact
arxiv_id, observed 2026-05-20T12:58:17.777018Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-20T12:56:48.177386Z digest=sha256:ae52d7f227c2ba1e64c687acb1ef8592f798eaa634e3ca1925d4abf17d109871

Observation da34c941-e438-4ce0-8660-5a33c323c6c7 · inbound

JuZhou 1.0 Technical Report: The First Edge-Native Text-to-Image Foundation Model Trained Entirely on China-Developed AI Accelerators cites this paper.

JuZhou 1.0 Technical Report: The First Edge-Native Text-to-Image Foundation Model Trained Entirely on China-Developed AI Accelerators Low-bit Model Quantization for Deep Neural Networks: A Survey

Reference 34

Resolution
metadata mismatch
arxiv_id, observed 2026-07-01T15:45:48.762968Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-06-30T01:02:20.024793Z digest=sha256:67ff2ce35976a431ccd22387d3dadfcc7cd2644f2e05be857e36fbb57dcdc3a3

Observation 498a9f54-498e-45fa-9745-2b9a6764cc15 · inbound

JuZhou 1.0 Technical Report: The First Edge-Native Text-to-Image Foundation Model Trained Entirely on China-Developed AI Accelerators cites this paper.

JuZhou 1.0 Technical Report: The First Edge-Native Text-to-Image Foundation Model Trained Entirely on China-Developed AI Accelerators Low-bit Model Quantization for Deep Neural Networks: A Survey

Reference 34

Resolution
unresolved
no resolver link, observed 2026-07-12T11:46:50.520083Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T11:46:50.520083Z digest=sha256:de67a37575029a1e1fc7bf60e6f4c1a9b83a4f9ca15b85e74a5fc1f9553f774c