Pith. sign in

Paper Citation Record · LEDGER

DynamicKV: Task-Aware Adaptive KV Cache Compression for Long Context LLMs

As of 5 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 15 inbound Pith citation observations for arXiv:2412.14838.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2412.14838 v4

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 15 of 15 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-04T06:34:03.388597+00:00

measured 15 of 15 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-03T03:16:19.908039Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

0
arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation ea16eb07-5299-44a1-9870-516b83da6ff3 · inbound

CompilerKV: Risk-Adaptive KV Compression via Offline Experience Compilation cites this paper.

CompilerKV: Risk-Adaptive KV Compression via Offline Experience Compilation DynamicKV: Task-Aware Adaptive KV Cache Compression for Long Context LLMs

Reference 20

Resolution
verified exact
arxiv_id, observed 2026-05-21T13:34:11.370570Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-21T13:32:15.598047Z digest=sha256:f300c388871f5a6a5329fe7acfb76dbf43b236157af3a0a9cc0005187a0f8ca4

Observation af94e715-099d-4c2d-a69c-569597dc15ef · inbound

CompilerKV: Risk-Adaptive KV Compression via Offline Experience Compilation cites this paper.

CompilerKV: Risk-Adaptive KV Compression via Offline Experience Compilation DynamicKV: Task-Aware Adaptive KV Cache Compression for Long Context LLMs

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-03T03:16:19.908039Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T03:16:19.908039Z digest=sha256:72b4318c1009dedf0ecd4ae8f325a65f574c76f41a48229a5139105d111c047e

Observation 3dee4f86-afda-4614-8bc4-3ce7403f73b3 · inbound

From Static Inference to Dynamic Interaction: A Survey of Streaming Large Language Models cites this paper.

From Static Inference to Dynamic Interaction: A Survey of Streaming Large Language Models DynamicKV: Task-Aware Adaptive KV Cache Compression for Long Context LLMs

Reference 14

Resolution
metadata mismatch
arxiv_id, observed 2026-05-15T16:20:09.970164Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-15T16:16:15.819622Z digest=sha256:69186f7c46e2a0916554bc025745429a2ed53523630e87a5bf94b8e60d5b2bfa

Observation fd5408ee-5436-4d65-8215-22112e7fc78e · inbound

Don't Waste Bits! Adaptive KV-Cache Quantization for Lightweight On-Device LLMs cites this paper.

Don't Waste Bits! Adaptive KV-Cache Quantization for Lightweight On-Device LLMs DynamicKV: Task-Aware Adaptive KV Cache Compression for Long Context LLMs

Reference 37

Resolution
verified exact
arxiv_id, observed 2026-05-10T22:55:48.123886Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-10T19:31:17.420639Z digest=sha256:8dd6e36868713b3f9a1c76269011cceec64142b2ae8c49ef005295a9c403eadd

Observation ea482e67-46c4-4806-bebd-45fefa3c12a7 · inbound

SAGE: Selective Attention-Guided Extraction for Token-Efficient Document Indexing cites this paper.

SAGE: Selective Attention-Guided Extraction for Token-Efficient Document Indexing DynamicKV: Task-Aware Adaptive KV Cache Compression for Long Context LLMs

Reference 55

Resolution
verified exact
arxiv_id, observed 2026-05-10T08:32:52.113523Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-10T08:32:02.222528Z digest=sha256:6c92ed76b9ae34bb83c9e7fc041add361f2d4d37bd56c3f7a33e77f03a6a93c4

Observation 91571be8-3c29-4401-8a00-d1c468f90051 · inbound

HieraSparse: Hierarchical Semi-Structured Sparse KV Attention cites this paper.

HieraSparse: Hierarchical Semi-Structured Sparse KV Attention DynamicKV: Task-Aware Adaptive KV Cache Compression for Long Context LLMs

Reference 28

Resolution
verified exact
arxiv_id, observed 2026-05-10T07:16:54.213976Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-10T07:15:19.184970Z digest=sha256:cc3e16a420b18dc7fd354a1a45fcb0b1dd008cb2f7131ab3e63fe334be7a0588

Observation 2bd6a4be-be7c-437f-a69b-bf517eedde51 · inbound

Network Edge Inference for Large Language Models: Principles, Techniques, and Opportunities cites this paper.

Network Edge Inference for Large Language Models: Principles, Techniques, and Opportunities DynamicKV: Task-Aware Adaptive KV Cache Compression for Long Context LLMs

Reference 211

Resolution
verified exact
arxiv_id, observed 2026-05-11T20:16:10.305729Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-08T09:45:57.201837Z digest=sha256:6d35d7f5d65ce8e9c4bdceea156e745078005d69354b5107283402e4473d6bfa

Observation 4ea9d3cb-195e-4348-b96e-e9815ba6a8d0 · inbound

ReST-KV: Robust KV Cache Eviction with Layer-wise Output Reconstruction and Spatial-Temporal Smoothing cites this paper.

ReST-KV: Robust KV Cache Eviction with Layer-wise Output Reconstruction and Spatial-Temporal Smoothing DynamicKV: Task-Aware Adaptive KV Cache Compression for Long Context LLMs

Reference 29

Resolution
verified exact
arxiv_id, observed 2026-05-12T07:26:30.323223Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-12T02:52:33.076123Z digest=sha256:4959c45bfff399317ef336fe4233d393d52f1bcd61fe98043ef76488da009d2f

Observation 2b30aa87-2826-4855-bf1a-f389a5bccd30 · inbound

YouZhi: Towards High-Concurrency Financial LLMs via Adaptive GQA-to-MLA Transition cites this paper.

YouZhi: Towards High-Concurrency Financial LLMs via Adaptive GQA-to-MLA Transition DynamicKV: Task-Aware Adaptive KV Cache Compression for Long Context LLMs

Reference 34

Resolution
metadata mismatch
arxiv_id, observed 2026-07-02T13:06:59.429187Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-06-28T01:33:54.383507Z digest=sha256:3773b4aba5f738ab5baf8e04bca8992cf2283ceb14ff95603730efb70b5abef7

Observation 2e9c34ce-26cb-40d9-90aa-e3a271609d4f · inbound

SpectrumKV: Per-Token Mixed-Precision KV Cache Transfer for Prefill-Decode Disaggregated LLM Serving cites this paper.

SpectrumKV: Per-Token Mixed-Precision KV Cache Transfer for Prefill-Decode Disaggregated LLM Serving DynamicKV: Task-Aware Adaptive KV Cache Compression for Long Context LLMs

Reference 22

Resolution
malformed identifier
arxiv_id, observed 2026-07-02T22:27:26.309060Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-06-27T18:50:50.066713Z digest=sha256:08938ca04c3d96f36fabc575e5c586e82a1937d469175d77802dd331c218be4e

Observation 5b49ead8-949a-43ce-a411-79add8557bad · inbound

IntentKV: Cross-Turn Intent-Aware KV Cache Pruning for Agent Inference cites this paper.

IntentKV: Cross-Turn Intent-Aware KV Cache Pruning for Agent Inference DynamicKV: Task-Aware Adaptive KV Cache Compression for Long Context LLMs

Reference 6

Resolution
metadata mismatch
arxiv_id, observed 2026-07-02T20:57:22.975598Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-06-27T20:04:44.703782Z digest=sha256:7a82c9dc19f3f528cd99f4149af22bbdd1bbd211701c066ee50aa6ed02ad2c5c

Observation fc489183-a0d4-4e26-94cd-a9ffddf66d45 · inbound

ReasonAlloc: Hierarchical Decoding-Time KV Cache Budget Allocation for Reasoning Models cites this paper.

ReasonAlloc: Hierarchical Decoding-Time KV Cache Budget Allocation for Reasoning Models DynamicKV: Task-Aware Adaptive KV Cache Compression for Long Context LLMs

Reference 14

Resolution
metadata mismatch
arxiv_id, observed 2026-07-03T05:37:40.355882Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-06-27T13:09:37.979051Z digest=sha256:ec5101dc04e1517c903d74339bb4183c4f7cf71d9df7bf1d4e6dfed5caadc4cd

Observation 2b1e6053-7804-403d-860e-cf5810fad4c0 · inbound

AnchorKV: Safety-Aware KV Cache Compression via Soft Penalty with a Refusal Anchor cites this paper.

AnchorKV: Safety-Aware KV Cache Compression via Soft Penalty with a Refusal Anchor DynamicKV: Task-Aware Adaptive KV Cache Compression for Long Context LLMs

Reference 11

Resolution
verified exact
arxiv_id, observed 2026-07-03T20:18:57.886130Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-06-27T01:23:23.488555Z digest=sha256:a7ed4501a68af4b2228de1de656fd1c06117778036626a202867dfae98ee75a8

Observation 04608168-8d5e-441b-8da9-e1fb4b2613fd · inbound

Finetuning Vision-Language-Action Models Requires Fewer Layers Than You Think cites this paper.

Finetuning Vision-Language-Action Models Requires Fewer Layers Than You Think DynamicKV: Task-Aware Adaptive KV Cache Compression for Long Context LLMs

Reference 32

Resolution
verified exact
arxiv_id, observed 2026-07-04T04:29:34.949974Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-06-26T16:59:28.243515Z digest=sha256:036b4485cc85475a3bbddf73d53a0462629ee75107988d030687227020e8b671

Observation a2b0cd05-6a7d-4054-aa8b-ca6dc843c250 · inbound

MemDecay: Region-Aware KV Cache Eviction for Efficient LLM Agent Inference cites this paper.

MemDecay: Region-Aware KV Cache Eviction for Efficient LLM Agent Inference DynamicKV: Task-Aware Adaptive KV Cache Compression for Long Context LLMs

Reference 29

Resolution
unresolved
no resolver link, observed 2026-07-14T10:41:08.967261Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T10:41:08.967261Z digest=sha256:568c673daf11d3df2372c002b86502c62c3e30a7b61431aac7aab4c4ff60ca9e