Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-02T21:53:20.715200Z
Paper Citation Record · LEDGER
As of 8 August 2026, this Paper Citation Record lists 13 of 13 outbound references and 0 inbound Pith citation observations for arXiv:2602.20191.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-02T21:53:20.715200Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
13 of 13 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 79cde160-0886-48e4-bcac-3135eb5ff04c · outbound
MoBiQuant: Mixture-of-Bits Quantization for Token-Adaptive Any-Precision LLM TruncQuant: Truncation-Ready Quantization for DNNs with Flexible Weight Bit Precision
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8d493c7c-a6a5-4df7-be60-42d327583dc7 · outbound
MoBiQuant: Mixture-of-Bits Quantization for Token-Adaptive Any-Precision LLM and Patterson, D
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8f5d2d71-e563-4c41-87f1-0f1877eabcef · outbound
MoBiQuant: Mixture-of-Bits Quantization for Token-Adaptive Any-Precision LLM J., and Lee, D
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0ec1bf84-1639-49c3-a580-a3077dd4f018 · outbound
MoBiQuant: Mixture-of-Bits Quantization for Token-Adaptive Any-Precision LLM Llama 2: Open Foundation and Fine-Tuned Chat Models
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3b04d30d-02c0-4e1c-85ff-71bbe0265205 · outbound
MoBiQuant: Mixture-of-Bits Quantization for Token-Adaptive Any-Precision LLM Overall Algorithm of MoBiQuant
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fad77491-c7aa-49b6-95d7-8cc9f00e1ec7 · outbound
MoBiQuant: Mixture-of-Bits Quantization for Token-Adaptive Any-Precision LLM LET preserves the main linear path by transforming the input activation and compensating the linear weights so that the layer output remains equivalent
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 86a81e30-791e-4a1e-8881-bd623c7df1f5 · outbound
MoBiQuant: Mixture-of-Bits Quantization for Token-Adaptive Any-Precision LLM Think you have Solved Question Answering? Try ARC, the AI2 Reasoning Challenge
Reference 2019
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bc863dbc-6cf7-4162-8af6-298115b03b33 · outbound
MoBiQuant: Mixture-of-Bits Quantization for Token-Adaptive Any-Precision LLM FGMP: Fine-Grained Mixed-Precision Weight and Activation Quantization for Hardware-Accelerated LLM Inference
Reference 2021
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f5730d99-596a-4f42-b972-8ba02549b4dd · outbound
MoBiQuant: Mixture-of-Bits Quantization for Token-Adaptive Any-Precision LLM GPTQ: Accurate Post-Training Quantization for Generative Pre-trained Transformers
Reference 2022
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 141b2ca0-3e07-4514-926d-badf21f465f3 · outbound
MoBiQuant: Mixture-of-Bits Quantization for Token-Adaptive Any-Precision LLM PrefixQuant: Eliminating Outliers by Prefixed Tokens for Large Language Models Quantization
Reference 2023
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 25bd78b3-a95f-442e-be59-9f3ed5ccdfc2 · outbound
MoBiQuant: Mixture-of-Bits Quantization for Token-Adaptive Any-Precision LLM A Survey of Quantization Methods for Efficient Neural Network Inference
Reference 2024
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 358df1f3-b77a-4a79-a336-cc2c15041569 · outbound
MoBiQuant: Mixture-of-Bits Quantization for Token-Adaptive Any-Precision LLM Mixtral of Experts
Reference 2025
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5668b894-29a7-45d9-a323-4b035854b1a3 · outbound
MoBiQuant: Mixture-of-Bits Quantization for Token-Adaptive Any-Precision LLM Pointer Sentinel Mixture Models
Reference 2026
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
No inbound Pith citation observations are available.