Pith. sign in

Paper Citation Record · LEDGER

KV Cache is 1 Bit Per Channel: Efficient Large Language Model Inference with Coupled Quantization

As of 9 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 4 inbound Pith citation observations for arXiv:2405.03917.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2405.03917 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 4 of 4 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 4 of 4 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-09T13:26:55.470592Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-07-10T01:36:44.144655Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 94352534-5c7c-4e04-bc9b-602f81faec4d · inbound

PolarQuant: Quantizing KV Caches with Polar Transformation cites this paper.

PolarQuant: Quantizing KV Caches with Polar Transformation KV Cache is 1 Bit Per Channel: Efficient Large Language Model Inference with Coupled Quantization

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-09T13:26:55.470592Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T13:26:55.470592Z digest=sha256:4843a9d3d0b150de027385b86ac93f340bf8ace0c18b658dae6f882b15846968

Observation 0c04dabc-3b5d-43dc-a42e-79054b65ac2a · inbound

TurboQuant: Online Vector Quantization with Near-optimal Distortion Rate cites this paper.

TurboQuant: Online Vector Quantization with Near-optimal Distortion Rate KV Cache is 1 Bit Per Channel: Efficient Large Language Model Inference with Coupled Quantization

Reference 65

Resolution
verified exact
arxiv_id, observed 2026-05-20T08:09:22.411208Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-20T08:09:22.226608Z digest=sha256:30649dfc6ad177b27cb1f3f1024e213c05a2c1a37b8cb5bd2e32a68d4314eb2b

Observation 0aef9934-a407-4c49-8162-428eb3976373 · inbound

CaliDrop: KV Cache Compression with Calibration cites this paper.

CaliDrop: KV Cache Compression with Calibration KV Cache is 1 Bit Per Channel: Efficient Large Language Model Inference with Coupled Quantization

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-06T13:57:15.821815Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:57:15.821815Z digest=sha256:75b175bd3d126bb89c34c15196c494ebfcbdd46b383b6eaf6390efe38f990ef0

Observation 5150f09a-df9f-40a6-a18e-a57a59854d2c · inbound

What to Keep, What to Forget: A Rate--Distortion View of Memory Compaction in LLMs and Agents cites this paper.

What to Keep, What to Forget: A Rate--Distortion View of Memory Compaction in LLMs and Agents KV Cache is 1 Bit Per Channel: Efficient Large Language Model Inference with Coupled Quantization

Reference 148

Resolution
verified exact
local_arxiv, observed 2026-07-10T01:36:44.146556Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-07-10T01:26:59.421158Z digest=sha256:ebc63b845be081c77b54a767bbf51decf60b43595af7d9ffb1d3fc1302a60e4b