Pith. sign in

Paper Citation Record · LEDGER

On the Efficacy of Eviction Policy for Key-Value Constrained Generative Language Model Inference

As of 17 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 11 inbound Pith citation observations for arXiv:2402.06262.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2402.06262 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 11 of 11 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-17T06:30:58.91139+00:00

measured 11 of 11 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-11T20:34:37.699018Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-07-10T00:56:40.724162Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation f0367217-df7e-48ff-b663-070b8cdaadd4 · inbound

Ada-KV: Optimizing KV Cache Eviction by Adaptive Budget Allocation for Efficient LLM Inference cites this paper.

Ada-KV: Optimizing KV Cache Eviction by Adaptive Budget Allocation for Efficient LLM Inference On the Efficacy of Eviction Policy for Key-Value Constrained Generative Language Model Inference

Reference 18

Resolution
verified exact
arxiv_id, observed 2026-05-17T11:16:32.074216Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-05-17T11:16:31.904921Z digest=sha256:b54f430dde6a45b4c280f4c87321bb29273256767f28171b360325806d81252f

Observation f3a5e6ca-a5ee-4335-86df-af172c5f7409 · inbound

Batch-Max: Higher LLM Throughput using Larger Batch Sizes and KV Cache Compression cites this paper.

Batch-Max: Higher LLM Throughput using Larger Batch Sizes and KV Cache Compression On the Efficacy of Eviction Policy for Key-Value Constrained Generative Language Model Inference

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-11T20:34:37.699018Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T20:34:37.699018Z digest=sha256:39db7feba9be038438bd92b8e0449008eb1ae3bbfa71ac2cf72bf4ddf4d78ba7

Observation 44e8655f-d43e-475f-9e8f-a1ed7056542a · inbound

ZigZagkv: Dynamic KV Cache Compression for Long-context Modeling based on Layer Uncertainty cites this paper.

ZigZagkv: Dynamic KV Cache Compression for Long-context Modeling based on Layer Uncertainty On the Efficacy of Eviction Policy for Key-Value Constrained Generative Language Model Inference

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-11T17:24:49.212636Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T17:24:49.212636Z digest=sha256:1d7ed2ee8af0e7924b35585736089d38bf2d2702a331bfc416e3ba0758ee2fb8

Observation 879af09d-5eb5-47d0-9bd8-6bd52689fa4d · inbound

More Tokens, Lower Precision: Towards the Optimal Token-Precision Trade-off in KV Cache Compression cites this paper.

More Tokens, Lower Precision: Towards the Optimal Token-Precision Trade-off in KV Cache Compression On the Efficacy of Eviction Policy for Key-Value Constrained Generative Language Model Inference

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-11T13:52:48.760345Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T13:52:48.760345Z digest=sha256:69868d929dd45cc0fd80c1c88b647bb51537f8a2af96cd642eec12f424a03877

Observation c21181a6-5cea-470f-a97a-05134b04e7ca · inbound

Memory-Efficient Visual Autoregressive Modeling with Scale-Aware KV Cache Compression cites this paper.

Memory-Efficient Visual Autoregressive Modeling with Scale-Aware KV Cache Compression On the Efficacy of Eviction Policy for Key-Value Constrained Generative Language Model Inference

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-07T14:16:17.384573Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:16:17.384573Z digest=sha256:5fff5130eb8572e92ace1325a3c596d0a90d960557dc38eefcd26298abdc6025

Observation a8e4ed20-fa5d-4f02-8769-64c78e86c8d7 · inbound

Meta-Soft: Leveraging Composable Meta-Tokens for Context-Preserving KV Cache Compression cites this paper.

Meta-Soft: Leveraging Composable Meta-Tokens for Context-Preserving KV Cache Compression On the Efficacy of Eviction Policy for Key-Value Constrained Generative Language Model Inference

Reference 22

Resolution
verified exact
arxiv_id, observed 2026-05-22T05:11:06.223367Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-05-22T05:10:53.133346Z digest=sha256:4a78dde5800752e284d0684df7ccc050c14b4b84ea227e26df4c55dbb017ef53

Observation db932869-3c78-43aa-9f62-0b8e715832b1 · inbound

Meta-Soft: Leveraging Composable Meta-Tokens for Context-Preserving KV Cache Compression cites this paper.

Meta-Soft: Leveraging Composable Meta-Tokens for Context-Preserving KV Cache Compression On the Efficacy of Eviction Policy for Key-Value Constrained Generative Language Model Inference

Reference 22

Resolution
verified exact
arxiv_id, observed 2026-06-30T17:34:57.862080Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-06-30T17:28:08.750049Z digest=sha256:b172b397116389bf93ec8bb5813c08ab7f9fa847793bfe68d33e8fbdb923e1e8

Observation dc14f934-a002-426d-b9a5-3a0d240b5ee9 · inbound

GRKV: Global Regression for Training-Free KV Cache Compression in Long-Context LLMs cites this paper.

GRKV: Global Regression for Training-Free KV Cache Compression in Long-Context LLMs On the Efficacy of Eviction Policy for Key-Value Constrained Generative Language Model Inference

Reference 23

Resolution
verified exact
arxiv_id, observed 2026-07-01T19:16:00.786285Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-06-28T22:54:55.101568Z digest=sha256:6e378842235fa1aab016bee674d6a21a94a352e90aef9db80278577f9657f48b

Observation 63b781bf-3a37-4545-a956-697178c20941 · inbound

HACK++: Towards More Effective Head-Aware Key-Value Compression for Efficient Visual Autoregressive Modeling cites this paper.

HACK++: Towards More Effective Head-Aware Key-Value Compression for Efficient Visual Autoregressive Modeling On the Efficacy of Eviction Policy for Key-Value Constrained Generative Language Model Inference

Reference 44

Resolution
verified exact
arxiv_id, observed 2026-07-02T21:27:24.332311Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-06-27T19:46:43.514413Z digest=sha256:ac76e7053d284ce1b92f9af1b44b9d40e2631100d653dfe5e8ef2d9cf948c7cc

Observation 1e580b97-9450-4d66-99fd-5732610d4ec0 · inbound

Coverage-Driven KV Cache Eviction for Efficient and Improved Inference of LLM cites this paper.

Coverage-Driven KV Cache Eviction for Efficient and Improved Inference of LLM On the Efficacy of Eviction Policy for Key-Value Constrained Generative Language Model Inference

Reference 77

Resolution
verified exact
arxiv_id, observed 2026-06-30T07:24:21.362514Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-06-30T07:19:48.530272Z digest=sha256:9e7389193d92c386624bc1b0bf4e85eb9c1af357c22528b7e6e09a82063ffa93

Observation 3c82e1a2-e934-47ec-9a94-dfe6b9bd0471 · inbound

Towards Efficient Large Language Model Serving: A Survey on System-Aware KV Cache Optimization cites this paper.

Towards Efficient Large Language Model Serving: A Survey on System-Aware KV Cache Optimization On the Efficacy of Eviction Policy for Key-Value Constrained Generative Language Model Inference

Reference 9

Resolution
malformed identifier
local_arxiv, observed 2026-07-10T00:56:40.725793Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-07-11T11:50:26.030339Z digest=sha256:9af0b08341ee2069d9bc3af10eb390bb063ec62b9ca1d478c2325a1277c7e702