Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 20 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 47 inbound Pith citation observations for arXiv:2410.09426.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-20T06:33:59.587034+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-15T22:32:56.673897Z
A source-named dated measurement, never combined with another source.
Source: pith, observed 2026-07-10T14:47:14.595381Z
0 of 0 outbound references displayed
External citation measurements
No source-named external measurement is stored.
No outbound reference observations are available for this paper version.
Observation cf63bef9-ed5c-4a67-ae3b-e16590eff196 · inbound
A Survey on Large Language Model Acceleration based on KV Cache Management FlatQuant: Flatness Matters for LLM Quantization
Reference 84
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 84f6cc41-98bc-4567-8db8-f5aa81e0f34e · inbound
MBQ: Modality-Balanced Quantization for Large Vision-Language Models FlatQuant: Flatness Matters for LLM Quantization
Reference 41
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation eb5a1c4f-ed55-40c2-817e-440bfcb4107f · inbound
Speculative Decoding Meets Quantization: Compatibility Evaluation and Hierarchical Framework Design FlatQuant: Flatness Matters for LLM Quantization
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f05e635b-5436-40eb-9de6-dfb5d3d91b5f · inbound
Turning LLM Activations Quantization-Friendly FlatQuant: Flatness Matters for LLM Quantization
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 252b32eb-f7f4-4bcd-809a-0b2ae9fc5dcc · inbound
FPTQuant: Function-Preserving Transforms for LLM Quantization FlatQuant: Flatness Matters for LLM Quantization
Reference 47
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9b4a16a2-bb2b-4aef-91ef-144e620cb39f · inbound
SmoothRot: Combining Channel-Wise Scaling and Rotation for Quantization-Friendly LLMs FlatQuant: Flatness Matters for LLM Quantization
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a2cf9096-26ba-445d-8fd4-16888dbb8644 · inbound
BTC-LLM: Efficient Sub-1-Bit LLM Quantization via Learnable Transformation and Binary Codebook FlatQuant: Flatness Matters for LLM Quantization
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation f7019349-af58-4c26-82d7-f4579f20ef7b · inbound
BASE-Q: Bias and Asymmetric Scaling Enhanced Rotational Quantization for Large Language Models FlatQuant: Flatness Matters for LLM Quantization
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 01d5bc78-5d95-421a-980a-e0fa4f1b8566 · inbound
Prune&Comp: Free Lunch for Layer-Pruned LLMs via Iterative Pruning with Magnitude Compensation FlatQuant: Flatness Matters for LLM Quantization
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ccd3cce7-639b-436b-b6df-46d6abd5dcd6 · inbound
OASIS: Outlier-Aware LUT-Based GEMM with Dual-Side Quantization for LLM Inference Acceleration FlatQuant: Flatness Matters for LLM Quantization
Reference 50
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1319fedc-277e-4ec4-ba9b-b618e122e71e · inbound
SQAP-VLA: A Synergistic Quantization-Aware Pruning Framework for High-Performance Vision-Language-Action Models FlatQuant: Flatness Matters for LLM Quantization
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ae612764-88b0-4a26-b26a-7d60345fe6c7 · inbound
ARCQuant: Boosting NVFP4 Quantization with Augmented Residual Channels for LLMs FlatQuant: Flatness Matters for LLM Quantization
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8a36bdb3-0c4d-42c5-9689-ceae74b09a9a · inbound
Pushing the Limits of Block Rotations in Post-Training Quantization FlatQuant: Flatness Matters for LLM Quantization
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 02f00097-e4a3-4008-a6d5-c9aede9f43db · inbound
Harmonia: Algorithm-Hardware Co-Design for Memory- and Compute-Efficient BFP-based LLM Inference FlatQuant: Flatness Matters for LLM Quantization
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 981620d0-e0a4-43c7-8720-32014b5ae3c4 · inbound
QuantVLA: Scale-Calibrated Post-Training Quantization for Vision-Language-Action Models FlatQuant: Flatness Matters for LLM Quantization
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation c7763bef-6454-406c-a39d-9a306fc8d6fa · inbound
RUQuant: Towards Refining Uniform Quantization for Large Language Models FlatQuant: Flatness Matters for LLM Quantization
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation dcfc9ff8-9d64-4e82-9ef8-749821b52979 · inbound
Efficient Matrix Implementation for Rotary Position Embedding FlatQuant: Flatness Matters for LLM Quantization
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation b0465a61-5f4c-4f65-acf2-80a765675ba0 · inbound
OSC: Hardware Efficient W4A4 Quantization via Outlier Separation in Channel Dimension FlatQuant: Flatness Matters for LLM Quantization
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 4816c971-df3b-45a9-b0d8-5af2f7d239fd · inbound
GSQ: Highly-Accurate Low-Precision Scalar Quantization for LLMs via Gumbel-Softmax Sampling FlatQuant: Flatness Matters for LLM Quantization
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 44500a70-ac63-4fce-ac0c-e5997af0ff5f · inbound
GSQ: Highly-Accurate Low-Precision Scalar Quantization for LLMs via Gumbel-Softmax Sampling FlatQuant: Flatness Matters for LLM Quantization
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 4b767354-3d73-497b-8c75-e9eaf5ebfc76 · inbound
QuantClaw: Precision Where It Matters for OpenClaw FlatQuant: Flatness Matters for LLM Quantization
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 5c406260-4f0a-4e6d-a86b-715e6a18e9a0 · inbound
TACO: Efficient Communication Compression of Intermediate Tensors for Scalable Tensor-Parallel LLM Training FlatQuant: Flatness Matters for LLM Quantization
Reference 49
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 509dd602-ecf0-4c42-9bb7-1ebd9f016c37 · inbound
GAMMA: Global Bit Allocation for Mixed-Precision Models under Arbitrary Budgets FlatQuant: Flatness Matters for LLM Quantization
Reference 55
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 07f6ca75-8adb-461f-8115-1407c96cc5e2 · inbound
Theory-optimal Quantization Based on Flatness FlatQuant: Flatness Matters for LLM Quantization
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 464745cb-6247-4472-931f-b022e77edc75 · inbound
Rotation-Aligned Key Channel Pruning for Efficient Vision-Language Model Inference FlatQuant: Flatness Matters for LLM Quantization
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 922a2180-cd71-4ba7-a745-bfda9b91541c · inbound
Breaking Modality Heterogeneity in Low-Bit Quantization for Large Vision-Language Models FlatQuant: Flatness Matters for LLM Quantization
Reference 42
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation a6059e17-cf3c-45de-bee1-2470236a5b27 · inbound
InfoQuant: Shaping Activation Distributions for Low-Bit LLM Quantization FlatQuant: Flatness Matters for LLM Quantization
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 0c447f2e-5f9b-43a5-bf3e-b11dbf452433 · inbound
HoloQ-VLA: Uniform W4A4 Quantization of Vision-Language-Action Models FlatQuant: Flatness Matters for LLM Quantization
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 01ce34e9-3f9d-40e5-83a8-6824d8a7311e · inbound
MixFP4: Enhancing NVFP4 with Adaptive FP4/INT4 Block Representations FlatQuant: Flatness Matters for LLM Quantization
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 281bc437-b119-420b-a0c2-c260d1492997 · inbound
Quantized Reasoning Models Think They Need to Think Longer, but They Do Not FlatQuant: Flatness Matters for LLM Quantization
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation de93df8b-96bc-4ebe-a4fa-0c1cdd2cd91f · inbound
Qift: Shift-Friendly No-Zero W2 Post-Training Quantization for Rotated W2A4/KV4 LLM Inference FlatQuant: Flatness Matters for LLM Quantization
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 90567447-2a14-48f3-b8c1-fc02e45bd621 · inbound
LiftQuant: Continuous Bit-Width LLM via Dimensional Lifting and Projection FlatQuant: Flatness Matters for LLM Quantization
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation d69724b8-b8fc-4f5c-afa9-14c835495fe9 · inbound
LiftQuant: Continuous Bit-Width LLM via Dimensional Lifting and Projection FlatQuant: Flatness Matters for LLM Quantization
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation b4439ba7-7369-45ef-99d1-1f41ec2b1ca1 · inbound
FAIR-Calib: Frontier-Aware Instability-Reweighted Calibration for Post-Training Quantization of Diffusion Large Language Models FlatQuant: Flatness Matters for LLM Quantization
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 452af151-a6dd-4d1e-9d41-a8134163b404 · inbound
FAIR-Calib: Frontier-Aware Instability-Reweighted Calibration for Post-Training Quantization of Diffusion Large Language Models FlatQuant: Flatness Matters for LLM Quantization
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 47f841d1-bf95-4a61-abce-ceaa57481538 · inbound
DynamicPTQ: Mitigating Activation Quantization Collapse via Residual-Stream Dynamics FlatQuant: Flatness Matters for LLM Quantization
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 87bfadae-b6c8-4e4a-bf82-93b4ea473d42 · inbound
HEPTv2: End-to-End Efficient Point Transformer for Charged Particle Reconstruction FlatQuant: Flatness Matters for LLM Quantization
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation feed6185-49ba-4e16-8cf6-4a52347fc662 · inbound
OrbitQuant: Data-Agnostic Quantization for Image and Video Diffusion Transformers FlatQuant: Flatness Matters for LLM Quantization
Reference 40
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation c2a1a963-30f2-49ff-a49d-f7efc7cd2cd0 · inbound
KronQ: LLM Quantization via Kronecker-Factored Hessian FlatQuant: Flatness Matters for LLM Quantization
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 23215029-05bd-423c-91ff-e93776deef77 · inbound
Quantize with Confidence? An Empirical Study of Quantization for Code Generation FlatQuant: Flatness Matters for LLM Quantization
Reference 53
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 90c9e76c-96a8-4dc8-862b-52a9df8e1451 · inbound
KroQuant: Kronecker-Structured Block Transforms for Efficient Post-Training Quantization of Diffusion Transformers FlatQuant: Flatness Matters for LLM Quantization
Reference 58
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d4ce794d-c345-4e23-878c-d5ed1a6d45c0 · inbound
GyRot: Leveraging Hidden Synergy between Rotation and Fine-grained Group Quantization for Low-bit LLM Inference FlatQuant: Flatness Matters for LLM Quantization
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 21d8ec9b-5d47-497b-8d34-32a4eedc50d2 · inbound
LightRot: A Light-Weighted Rotation Scheme and Architecture for Accurate Low-Bit Large Language Model Inference FlatQuant: Flatness Matters for LLM Quantization
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b1728799-5ba9-4fb7-937e-88df2418a78c · inbound
Hidden Language Consistency Phenomena in Reasoning LLMs FlatQuant: Flatness Matters for LLM Quantization
Reference 248
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation dc904aee-37cb-4221-ab88-8c1eb790edd5 · inbound
Understanding Calibration and Truncation Error Propagation in Training-Free Low-Rank Compression for LLMs FlatQuant: Flatness Matters for LLM Quantization
Reference 108
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d368b82d-868b-453d-8ba0-de167cbda201 · inbound
UnionSparse: An Index-Efficient Sparsity Framework for Low-Bit Sparse LLM Inference on Edge FlatQuant: Flatness Matters for LLM Quantization
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4b7a28b2-59b1-4ad8-a9c2-401784380ae3 · inbound
When Local Variance Optimality Is Not Enough: RoPE-Aligned Q/K Rotations for Dynamic 4-Bit Quantisation FlatQuant: Flatness Matters for LLM Quantization
Reference 42
Source-reported events for the cited work
Unavailable: canonical work link unavailable.