Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 7 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 12 inbound Pith citation observations for arXiv:2406.07056.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-07T13:32:34.524396Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-06-28T19:12:34.646797Z
0 of 0 outbound references displayed
External citation measurements
No source-named external measurement is stored.
No outbound reference observations are available for this paper version.
Observation ca559c28-8d52-4068-a408-0d456a2caf2c · inbound
LightTransfer: Your Long-Context LLM is Secretly a Hybrid Model with Effortless Adaptation Effectively Compress KV Heads for LLM
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation d5467938-fa24-4ea2-884c-4e003de5fc59 · inbound
Hardware-Efficient Attention for Fast Decoding Effectively Compress KV Heads for LLM
Reference 78
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0f93231f-9814-4dc5-85c4-ecb43a2b7e43 · inbound
TaDA: Training-free recipe for Decoding with Adaptive KV Cache Compression and Mean-centering Effectively Compress KV Heads for LLM
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1ce2cfb8-7c12-45f3-b120-633a5b4dde42 · inbound
Cartridges: Lightweight and general-purpose long context representations via self-study Effectively Compress KV Heads for LLM
Reference 105
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 59e4af24-f7c0-4ad3-8846-51f8fbccd7e4 · inbound
KV-Latent: Dimensional-level KV Cache Reduction with Frequency-aware Rotary Positional Embedding Effectively Compress KV Heads for LLM
Reference 46
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2ccd4bf6-e3fa-41e6-823b-ffa68218497f · inbound
LaCache: Ladder-Shaped KV Caching for Efficient Long-Context Modeling of Large Language Models Effectively Compress KV Heads for LLM
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3d124d0f-0688-4e16-8aea-961b5b574e41 · inbound
CaliDrop: KV Cache Compression with Calibration Effectively Compress KV Heads for LLM
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 129d056b-aa76-4449-8ef0-fd01fe67112e · inbound
TPLA: Tensor Parallel Latent Attention for Efficient Disaggregated Prefill and Decode Inference Effectively Compress KV Heads for LLM
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6d4fd0d3-0437-4586-b14c-22365ba354aa · inbound
WSVD: Weighted Low-Rank Approximation for Fast and Efficient Execution of Low-Precision Vision-Language Models Effectively Compress KV Heads for LLM
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 88fd4ed8-96dd-4c18-b2c2-ce95482be7a3 · inbound
DREAM-S: Speculative Decoding with Searchable Drafting and Target-Aware Refinement for Multimodal Generation Effectively Compress KV Heads for LLM
Reference 58
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation adf25fc3-b10d-4806-a842-42e91bdda4b0 · inbound
LASER: Loss-Aware Singular-value Decomposition and Rank Allocation for Efficient Low-Precision Vision-Language Models Effectively Compress KV Heads for LLM
Reference 55
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 906ecff6-6869-4d8a-9b1c-114a7c9ddbb5 · inbound
SOS-LoRA: Static Orthogonal-Subspace Low-Rank Adaptation with Fixed Multi-Scale Scaling Effectively Compress KV Heads for LLM
Reference 47
Source-reported events for the cited work
Unavailable: canonical work link unavailable.