Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-16T05:52:21.268115Z
Paper Citation Record · LEDGER
As of 23 August 2026, this Paper Citation Record lists 29 of 29 outbound references and 0 inbound Pith citation observations for arXiv:2504.19746.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-16T05:52:21.268115Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-23T06:30:58.430688+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
29 of 29 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 6751774a-8f98-4987-9cbb-8bf2f4250342 · outbound
FineQ: Software-Hardware Co-Design for Low-Bit Fine-Grained Mixed-Precision Quantization of LLMs Attention is all you need,
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 638ce25d-5a80-4655-8c64-39e2374a410a · outbound
FineQ: Software-Hardware Co-Design for Low-Bit Fine-Grained Mixed-Precision Quantization of LLMs LLaMA: Open and Efficient Foundation Language Models
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 05e56dfc-d151-491d-af8e-5d6043306928 · outbound
FineQ: Software-Hardware Co-Design for Low-Bit Fine-Grained Mixed-Precision Quantization of LLMs ubrain: A unary brain computer interface,
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation c83e3c9b-cc66-4d61-868d-0b9ebd906c2d · outbound
FineQ: Software-Hardware Co-Design for Low-Bit Fine-Grained Mixed-Precision Quantization of LLMs Up or down? adaptive rounding for post-training quantization,
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation eb4f35e2-de4f-447a-8f54-c965a1f6daab · outbound
FineQ: Software-Hardware Co-Design for Low-Bit Fine-Grained Mixed-Precision Quantization of LLMs GPTQ: Accurate Post-Training Quantization for Generative Pre-trained Transformers
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4456aaf4-3e8a-4940-a502-aa661fe3f3ab · outbound
FineQ: Software-Hardware Co-Design for Low-Bit Fine-Grained Mixed-Precision Quantization of LLMs Zeroquant: Efficient and affordable post-training quantization for large- scale transformers,
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9c7da2ac-8051-4cb3-85b5-2855d55cd0bd · outbound
FineQ: Software-Hardware Co-Design for Low-Bit Fine-Grained Mixed-Precision Quantization of LLMs LLM.int8(): 8-bit Matrix Multiplication for Transformers at Scale
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f608a497-daff-4357-a017-dcb07b099ff8 · outbound
FineQ: Software-Hardware Co-Design for Low-Bit Fine-Grained Mixed-Precision Quantization of LLMs Awq: Activation-aware weight quanti- zation for on-device llm compression and acceleration,
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 06a58058-8e0c-40d1-9a29-f14523a891d8 · outbound
FineQ: Software-Hardware Co-Design for Low-Bit Fine-Grained Mixed-Precision Quantization of LLMs Owq: Outlier-aware weight quantization for efficient fine-tuning and inference of large language models,
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 62505196-e02f-4ea2-808b-c0f77d7744ec · outbound
FineQ: Software-Hardware Co-Design for Low-Bit Fine-Grained Mixed-Precision Quantization of LLMs Llm-mq: Mixed-precision quantization for efficient llm deployment,
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation fc1bc929-847a-484a-a88b-f37e9b9c8f78 · outbound
FineQ: Software-Hardware Co-Design for Low-Bit Fine-Grained Mixed-Precision Quantization of LLMs PB-LLM: Partially Binarized Large Language Models
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3527ea1e-b722-4025-9e9b-c938f18f1eeb · outbound
FineQ: Software-Hardware Co-Design for Low-Bit Fine-Grained Mixed-Precision Quantization of LLMs Olive: Accelerating large language models via hardware- friendly outlier-victim pair quantization,
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9f2d7edb-3921-4959-865f-af6472d1851e · outbound
FineQ: Software-Hardware Co-Design for Low-Bit Fine-Grained Mixed-Precision Quantization of LLMs Tender: Accelerating Large Language Models via Tensor Decomposition and Runtime Requantization
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4811a75c-1c52-4c44-9085-3a46c9967802 · outbound
FineQ: Software-Hardware Co-Design for Low-Bit Fine-Grained Mixed-Precision Quantization of LLMs Gpt3. int8 (): 8-bit matrix multiplication for transformers at scale,
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e28c5483-fe6d-4372-bfa6-545fc1be7c74 · outbound
FineQ: Software-Hardware Co-Design for Low-Bit Fine-Grained Mixed-Precision Quantization of LLMs Abstractive long text summarization using large language models,
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 3ef05d73-7f2b-4db3-ad4c-49de2d765f03 · outbound
FineQ: Software-Hardware Co-Design for Low-Bit Fine-Grained Mixed-Precision Quantization of LLMs Transformer models used for text-based question answering systems,
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 438f8120-5cef-4a00-9b57-a056c1abd8b6 · outbound
FineQ: Software-Hardware Co-Design for Low-Bit Fine-Grained Mixed-Precision Quantization of LLMs Efficient memory management for large language model serving with pagedattention,
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b2296042-e679-4ffd-8581-3eab77aea4a8 · outbound
FineQ: Software-Hardware Co-Design for Low-Bit Fine-Grained Mixed-Precision Quantization of LLMs Pointer Sentinel Mixture Models
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 429d3f48-cfa7-43a3-a3ee-10f6c5f1f093 · outbound
FineQ: Software-Hardware Co-Design for Low-Bit Fine-Grained Mixed-Precision Quantization of LLMs Exploring the limits of transfer learning with a unified text-to-text transformer,
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c9b1d251-a023-4c40-8716-8f7cb0d88ab0 · outbound
FineQ: Software-Hardware Co-Design for Low-Bit Fine-Grained Mixed-Precision Quantization of LLMs Kurup and T
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 7626c832-3930-4635-975f-c69ffffbbc7a · outbound
FineQ: Software-Hardware Co-Design for Low-Bit Fine-Grained Mixed-Precision Quantization of LLMs Ascend-freepdk45: An open source standard cell library for asyn- chronous design,
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 4122b31b-5767-496a-92b4-3d0843d4a72d · outbound
FineQ: Software-Hardware Co-Design for Low-Bit Fine-Grained Mixed-Precision Quantization of LLMs Efficient processing of deep neural networks: A tutorial and survey,
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 74ce057f-56fe-4998-8e35-af6173c88172 · outbound
FineQ: Software-Hardware Co-Design for Low-Bit Fine-Grained Mixed-Precision Quantization of LLMs A Survey of Large Language Models
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8c4c853e-fc2a-4b1b-b7df-37648d77476c · outbound
FineQ: Software-Hardware Co-Design for Low-Bit Fine-Grained Mixed-Precision Quantization of LLMs SqueezeLLM: Dense-and-Sparse Quantization
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b972c15e-35e2-4a06-877e-6224ca043421 · outbound
FineQ: Software-Hardware Co-Design for Low-Bit Fine-Grained Mixed-Precision Quantization of LLMs APTQ: Attention-aware Post-Training Mixed-Precision Quantization for Large Language Models
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7674783b-f7ee-4006-b928-3368671632ba · outbound
FineQ: Software-Hardware Co-Design for Low-Bit Fine-Grained Mixed-Precision Quantization of LLMs Understanding reuse, performance, and hardware cost of dnn dataflow: A data-centric approach,
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4a266efd-cef3-40aa-9bef-13ecc57e9ac1 · outbound
FineQ: Software-Hardware Co-Design for Low-Bit Fine-Grained Mixed-Precision Quantization of LLMs Outlier Suppression+: Accurate quantization of large language models by equivalent and optimal shifting and scaling
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 25103a6c-4a20-48cc-9945-ffff86ceb0a7 · outbound
FineQ: Software-Hardware Co-Design for Low-Bit Fine-Grained Mixed-Precision Quantization of LLMs Mobile and edge evaluation of large language models,
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 8fbfc193-ed8d-4c28-bc7c-1808fd5cedab · outbound
FineQ: Software-Hardware Co-Design for Low-Bit Fine-Grained Mixed-Precision Quantization of LLMs Hardware-aware parallel prompt decoding for memory-efficient acceleration of llm inference,
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
No inbound Pith citation observations are available.