Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-10T16:39:33.130911Z
Paper Citation Record · LEDGER
As of 11 August 2026, this Paper Citation Record lists 23 of 23 outbound references and 0 inbound Pith citation observations for arXiv:2501.12956.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-10T16:39:33.130911Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-11T06:34:44.6726+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
23 of 23 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 2c517ffd-14d7-4958-8c9b-a7514b41cea5 · outbound
GANQ: GPU-Adaptive Non-Uniform Quantization for Large Language Models GPT-4 Technical Report
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation aa2789a8-956f-424c-b61b-ee1253462a75 · outbound
GANQ: GPU-Adaptive Non-Uniform Quantization for Large Language Models The Llama 3 Herd of Models
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ae6a6cfe-e605-47a4-8ad3-6bef7fe0b62b · outbound
GANQ: GPU-Adaptive Non-Uniform Quantization for Large Language Models Gemini: A Family of Highly Capable Multimodal Models
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e1a5d41e-b01d-4b3e-85eb-d8f6afae6bbc · outbound
GANQ: GPU-Adaptive Non-Uniform Quantization for Large Language Models Fast matrix multiplications for lookup table-quantized llms
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 8bc6ef43-27ae-4835-8f6b-7ac349944999 · outbound
GANQ: GPU-Adaptive Non-Uniform Quantization for Large Language Models Scaling Laws for Neural Language Models
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2ea6260f-dbee-4083-9a4d-6e25561ee95c · outbound
GANQ: GPU-Adaptive Non-Uniform Quantization for Large Language Models X., Nie, J.-Y ., and Wen, J.-R
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1b6b94ea-19b0-4aca-8210-3fec4b411edd · outbound
GANQ: GPU-Adaptive Non-Uniform Quantization for Large Language Models Llm-qat: Data-free quantization aware training for large language models
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 5826ddfa-8a06-47a9-9793-ec4bda5a408b · outbound
GANQ: GPU-Adaptive Non-Uniform Quantization for Large Language Models A., MacIntyre, R., Bies, A., Ferguson, M., Katz, K., and Schasberger, B
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation a4025208-1a45-4d00-bec7-8d4464fd1dad · outbound
GANQ: GPU-Adaptive Non-Uniform Quantization for Large Language Models LLaMA: Open and Efficient Foundation Language Models
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 814adf65-71dd-4d0a-b1c4-b592457988cd · outbound
GANQ: GPU-Adaptive Non-Uniform Quantization for Large Language Models Augmenting Black-box LLMs with Medical Textbooks for Biomedical Question Answering
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fe1ec67c-a8aa-4c2c-99f5-bdbbcfd8c571 · outbound
GANQ: GPU-Adaptive Non-Uniform Quantization for Large Language Models HuggingFace's Transformers: State-of-the-art Natural Language Processing
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e18376b8-e0dd-4e98-8c91-0559c3b0abc2 · outbound
GANQ: GPU-Adaptive Non-Uniform Quantization for Large Language Models HellaSwag: Can a Machine Really Finish Your Sentence?
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 42c6e2f1-8283-46da-8389-668bb169e234 · outbound
GANQ: GPU-Adaptive Non-Uniform Quantization for Large Language Models OPT: Open Pre-trained Transformer Language Models
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6121e131-c2a5-4321-9291-395e8aa041ae · outbound
GANQ: GPU-Adaptive Non-Uniform Quantization for Large Language Models Unresolved cited work
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 4d280e49-f035-4377-9229-8007e5dd0d6a · outbound
GANQ: GPU-Adaptive Non-Uniform Quantization for Large Language Models GPTQ: Accurate Post-Training Quantization for Generative Pre-trained Transformers
Reference 1925
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e5befff7-2b77-46eb-9120-8a117f26f522 · outbound
GANQ: GPU-Adaptive Non-Uniform Quantization for Large Language Models LUT Tensor Core: A Software-Hardware Co-Design for LUT-Based Low-Bit LLM Inference
Reference 2017
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8e3cd454-0649-4207-a733-76ab915a6aa4 · outbound
GANQ: GPU-Adaptive Non-Uniform Quantization for Large Language Models Training Verifiers to Solve Math Word Problems
Reference 2018
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation aa5f5249-3049-4a86-9084-2bacfb5e58a6 · outbound
GANQ: GPU-Adaptive Non-Uniform Quantization for Large Language Models Think you have Solved Question Answering? Try ARC, the AI2 Reasoning Challenge
Reference 2019
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7d3debeb-5733-4a17-a01c-95f6d60c1710 · outbound
GANQ: GPU-Adaptive Non-Uniform Quantization for Large Language Models SpQR: A Sparse-Quantized Representation for Near-Lossless LLM Weight Compression
Reference 2021
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6a79fdb1-6923-4457-901b-1c455ce25e1e · outbound
GANQ: GPU-Adaptive Non-Uniform Quantization for Large Language Models NUPES : Non-Uniform Post-Training Quantization via Power Exponent Search
Reference 2022
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f582e29a-cd3e-4074-bc11-45edd6b1c7bb · outbound
GANQ: GPU-Adaptive Non-Uniform Quantization for Large Language Models Boolq: Exploring the surprising difficulty of natural yes/no questions
Reference 2023
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation e8e4dac9-0063-4553-a53f-0c832cb86a8c · outbound
GANQ: GPU-Adaptive Non-Uniform Quantization for Large Language Models D., Dhariwal, P., Neelakantan, A., Shyam, P., Sastry, G., Askell, A., et al
Reference 2024
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d8f54638-188a-4ceb-a98f-3d4c61185778 · outbound
GANQ: GPU-Adaptive Non-Uniform Quantization for Large Language Models Deep Compression: Compressing Deep Neural Networks with Pruning, Trained Quantization and Huffman Coding
Reference 2025
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
No inbound Pith citation observations are available.