Pith. sign in

Paper Citation Record · LEDGER

EvolKV: Evolutionary KV Cache Compression for LLM Inference

As of 10 August 2026, this Paper Citation Record lists 54 of 54 outbound references and 0 inbound Pith citation observations for arXiv:2509.08315.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2509.08315 v1

Coverage vector

measured 54 of 54 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-04T20:53:30.792140Z

measured 54 of 54 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

54 of 54 outbound references displayed

  • verified exact7
  • verified fuzzy1
  • unresolved46
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 4e95db61-9d48-40b4-87ce-6ed386ad51f7 · outbound

This paper cites URL: " 'urlintro :=.

EvolKV: Evolutionary KV Cache Compression for LLM Inference URL: " 'urlintro :=

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-04T20:53:26.291319Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T20:53:26.291319Z digest=sha256:6b5a6e2c9f7149a1768857e322cdb40b2f80a3bd70b6ce6a3288d038b87d5110

Observation 038f5414-d74b-428b-88ff-b2c5ea509870 · outbound

This paper cites write newline.

EvolKV: Evolutionary KV Cache Compression for LLM Inference write newline

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-04T20:53:26.358653Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T20:53:26.358653Z digest=sha256:23d114dcde15c84fb728db6cc28e8e26bf8490ede6f886daf17f3e511dc1a13e

Observation 9e27fdca-227a-41f1-882f-7a0ad92a7a99 · outbound

This paper cites Genetic Algorithm: Reviews, Implementations, and Applications.

EvolKV: Evolutionary KV Cache Compression for LLM Inference Genetic Algorithm: Reviews, Implementations, and Applications

Reference 3

Resolution
verified exact
local_arxiv, observed 2026-08-04T20:53:32.231576Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-04T20:53:26.518273Z digest=sha256:003b260e6fd5de9ddb84ced8c229821bd173091c6d48ef18ef00e58a5429fac4

Observation 593b2dde-4a6e-472f-9c65-39b467271ec6 · outbound

This paper cites Qwen Technical Report.

EvolKV: Evolutionary KV Cache Compression for LLM Inference Qwen Technical Report

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-04T20:53:26.644643Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T20:53:26.644643Z digest=sha256:2ec77c65ad4ab935640d1261375416f6089961d7f13c1915444c2b42605ce6f7

Observation 74b606d7-a608-4507-9008-6a668a2b8a28 · outbound

This paper cites LongBench: A Bilingual, Multitask Benchmark for Long Context Understanding.

EvolKV: Evolutionary KV Cache Compression for LLM Inference LongBench: A Bilingual, Multitask Benchmark for Long Context Understanding

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-04T20:53:26.772769Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T20:53:26.772769Z digest=sha256:d8801d97bc3cad771b9fde851233e18bf5f0ae456e5cbf82397d7e4fa79fb2a5

Observation 806b7f12-12a1-4b0f-a31a-73fabaec4ad6 · outbound

This paper cites an unresolved cited work.

EvolKV: Evolutionary KV Cache Compression for LLM Inference Unresolved cited work

Reference 6

Resolution
verified exact
doi, observed 2026-08-04T20:53:31.588316Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-04T20:53:26.981819Z digest=sha256:75d3361ce5b3fc2337faa76594d69487bf5c3c7341b04ae91131854e3c39d2b5

Observation 37e36d34-34d6-458f-b35b-8db03f0729a8 · outbound

This paper cites Longformer: The Long-Document Transformer.

EvolKV: Evolutionary KV Cache Compression for LLM Inference Longformer: The Long-Document Transformer

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-04T20:53:27.205605Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T20:53:27.205605Z digest=sha256:b6d8e9558931b782f2eca4fea998001a546b7fa7035ad65d2647fa5c58e4a894

Observation ec5e6bb3-5d9a-4eb8-be3f-f54dad7cc057 · outbound

This paper cites PyramidKV: Dynamic KV Cache Compression based on Pyramidal Information Funneling.

EvolKV: Evolutionary KV Cache Compression for LLM Inference PyramidKV: Dynamic KV Cache Compression based on Pyramidal Information Funneling

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-04T20:53:27.311694Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T20:53:27.311694Z digest=sha256:c897daa254cc24b65dfd1707c7e3784ab3fa27a2b829b725ca0eeae6d69a0851

Observation 8d1def59-fa32-485f-b2c1-cfb3e04dcd89 · outbound

This paper cites an unresolved cited work.

EvolKV: Evolutionary KV Cache Compression for LLM Inference Unresolved cited work

Reference 9

Resolution
verified exact
doi, observed 2026-08-04T20:53:31.437565Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-04T20:53:27.372470Z digest=sha256:5217cab83438aa1214c747d51b60aa9e709939114e8ae84b1ab72c7eb4aa0421

Observation 9ef49c92-b34f-4faa-9cfa-c4499f9355f9 · outbound

This paper cites an unresolved cited work.

EvolKV: Evolutionary KV Cache Compression for LLM Inference Unresolved cited work

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-04T20:53:27.452212Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T20:53:27.452212Z digest=sha256:6f87cf53385107b965cc52df6a435e3e50c44f18e203184f127934f01750ef80

Observation a2b9a39e-aff6-4078-bc8b-a51f1dffcb31 · outbound

This paper cites an unresolved cited work.

EvolKV: Evolutionary KV Cache Compression for LLM Inference Unresolved cited work

Reference 11

Resolution
verified exact
doi, observed 2026-08-04T20:53:31.302934Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-04T20:53:27.526927Z digest=sha256:3fab2281b7cce0e8c2d0db34ef2dc8ee614668482570e7bd9dc17ef0cff7dcd2

Observation 227e76f9-5a1e-42ce-a729-3f38e495cddf · outbound

This paper cites an unresolved cited work.

EvolKV: Evolutionary KV Cache Compression for LLM Inference Unresolved cited work

Reference 12

Resolution
verified exact
doi, observed 2026-08-04T20:53:31.172668Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-04T20:53:27.618861Z digest=sha256:8537631980df31cde0011d6ef3260bb17a44b346612e6cb8f53af3f50a674c2b

Observation dc81beab-0131-4406-a83d-f52ac40a97be · outbound

This paper cites Generating Long Sequences with Sparse Transformers.

EvolKV: Evolutionary KV Cache Compression for LLM Inference Generating Long Sequences with Sparse Transformers

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-04T20:53:27.664702Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T20:53:27.664702Z digest=sha256:d2fc8afe2462c950864a1a0e00a9a2250419c6ed016b7ef6a77f4e8713794642

Observation b2a0d275-84b9-494b-a5c7-97df1e1d456b · outbound

This paper cites Training Verifiers to Solve Math Word Problems.

EvolKV: Evolutionary KV Cache Compression for LLM Inference Training Verifiers to Solve Math Word Problems

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-04T20:53:27.718150Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T20:53:27.718150Z digest=sha256:4ee259beecc4d50e28c6eccdfa22e0075281bfa8f9701da7d9b43bb55f34a3f1

Observation b2669a8f-751b-4fc0-8631-0cfca693f10f · outbound

This paper cites FlashAttention: Fast and Memory-Efficient Exact Attention with IO-Awareness.

EvolKV: Evolutionary KV Cache Compression for LLM Inference FlashAttention: Fast and Memory-Efficient Exact Attention with IO-Awareness

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-04T20:53:27.863745Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T20:53:27.863745Z digest=sha256:2485afee693daef91ae99f9d5963cef836102e6d49073e8ac9dacc8e006000a0

Observation 97110207-a1d0-4143-a3e6-5a3e99adea6f · outbound

This paper cites A Dataset of Information-Seeking Questions and Answers Anchored in Research Papers.

EvolKV: Evolutionary KV Cache Compression for LLM Inference A Dataset of Information-Seeking Questions and Answers Anchored in Research Papers

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-04T20:53:27.975894Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T20:53:27.975894Z digest=sha256:d7dcdbcacc164e5139b26124fa863e51d41b1db5771d4a68f5c619b58ff5adf3

Observation d479713b-ab9a-4d4a-af37-fc942267fce2 · outbound

This paper cites an unresolved cited work.

EvolKV: Evolutionary KV Cache Compression for LLM Inference Unresolved cited work

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-04T20:53:28.025096Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T20:53:28.025096Z digest=sha256:2f6b24ff92c6f65682587c00c21f5c49748d025dfaf708dd62c7984ba98d43e0

Observation 89d44375-6ca3-4c16-b863-3d71f108746a · outbound

This paper cites Model Tells You What to Discard: Adaptive KV Cache Compression for LLMs.

EvolKV: Evolutionary KV Cache Compression for LLM Inference Model Tells You What to Discard: Adaptive KV Cache Compression for LLMs

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-04T20:53:28.117122Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T20:53:28.117122Z digest=sha256:acee69ce09069cc3c25cfa37334cc372475de3d8a1e244f00f2d2294b754ce12

Observation ffd10026-e907-405f-b4ef-93185d0e7f7b · outbound

This paper cites an unresolved cited work.

EvolKV: Evolutionary KV Cache Compression for LLM Inference Unresolved cited work

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-04T20:53:28.177042Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T20:53:28.177042Z digest=sha256:a28eaf112cbc4ca5b081e329924332bb5c72b0cdfff542e11865b5a2e21d7395

Observation 25f304d6-cb41-4a7e-b292-c6c5c8907899 · outbound

This paper cites The Llama 3 Herd of Models.

EvolKV: Evolutionary KV Cache Compression for LLM Inference The Llama 3 Herd of Models

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-04T20:53:28.280449Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T20:53:28.280449Z digest=sha256:4defbd3b1fdb7a527093405abd6989dfb2339748b4c73254c3628f4ecc6b58bc

Observation cce07c45-05bf-4860-8809-92152efa642c · outbound

This paper cites LongCoder: A Long-Range Pre-trained Language Model for Code Completion.

EvolKV: Evolutionary KV Cache Compression for LLM Inference LongCoder: A Long-Range Pre-trained Language Model for Code Completion

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-04T20:53:28.374598Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T20:53:28.374598Z digest=sha256:2efbad15f1225ddd0d05fb7167cb329fe1feaea47a6d8c77f62a55927857321c

Observation c7da493d-c064-4e38-aab2-bb054bd88822 · outbound

This paper cites Müller, and Petros Koumoutsakos.

EvolKV: Evolutionary KV Cache Compression for LLM Inference Müller, and Petros Koumoutsakos

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-04T20:53:28.492763Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T20:53:28.492763Z digest=sha256:03255e7594b95f6b5ce3d35eab942120d4b69241044b69c0d8de50e9959db5b2

Observation 88d67d04-10e9-4b9e-91fb-5cbfd1af2c05 · outbound

This paper cites Constructing A Multi-hop QA Dataset for Comprehensive Evaluation of Reasoning Steps.

EvolKV: Evolutionary KV Cache Compression for LLM Inference Constructing A Multi-hop QA Dataset for Comprehensive Evaluation of Reasoning Steps

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-04T20:53:28.604509Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T20:53:28.604509Z digest=sha256:d98cf751ae6c9b359f5b0dfcff979d837591187957f575f15f47a5ce329bd750

Observation 02bce1f9-b549-4454-89d1-0fae7b300aff · outbound

This paper cites an unresolved cited work.

EvolKV: Evolutionary KV Cache Compression for LLM Inference Unresolved cited work

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-04T20:53:28.708057Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T20:53:28.708057Z digest=sha256:2672e94288b9cb533f0c84fc3cb469d70c0614a3b549aea8eceffd6f806c0de7

Observation 2aa56bfe-9d3d-4cfb-8a1b-2fa388d93e0b · outbound

This paper cites RULER: What's the Real Context Size of Your Long-Context Language Models?.

EvolKV: Evolutionary KV Cache Compression for LLM Inference RULER: What's the Real Context Size of Your Long-Context Language Models?

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-04T20:53:28.828244Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T20:53:28.828244Z digest=sha256:15367dcb156db44574f4144c1ca6ea9c005eaae012d58fedb63a8d3f3c62408c

Observation dd278312-35d1-4ccc-a742-c2ccea7cd444 · outbound

This paper cites an unresolved cited work.

EvolKV: Evolutionary KV Cache Compression for LLM Inference Unresolved cited work

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-04T20:53:29.006808Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T20:53:29.006808Z digest=sha256:467ab48ba3810a2dd89220f95461cf3eb26494e9e095365896e0245604a0c012

Observation 3cf5adb1-59b7-46cb-b6ee-be469b6c2134 · outbound

This paper cites Efficient Attentions for Long Document Summarization.

EvolKV: Evolutionary KV Cache Compression for LLM Inference Efficient Attentions for Long Document Summarization

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-04T20:53:29.091290Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T20:53:29.091290Z digest=sha256:c0d1d3973a12ea578b674c79d4206616fd9b4cfab0ba058a52f1a815ce2e118c

Observation 1f9b3a52-e930-4d2c-88a6-387b61f4d6e5 · outbound

This paper cites Mistral 7B.

EvolKV: Evolutionary KV Cache Compression for LLM Inference Mistral 7B

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-04T20:53:29.222206Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T20:53:29.222206Z digest=sha256:112ab218d170f326673cfa3d5bca99715a45597a7aae37639d7ac81677ee1a67

Observation 3aff6643-38d1-4803-a35e-cdf8c451fdf9 · outbound

This paper cites an unresolved cited work.

EvolKV: Evolutionary KV Cache Compression for LLM Inference Unresolved cited work

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-04T20:53:29.329672Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T20:53:29.329672Z digest=sha256:9685068ef2e10372292ae0bb08d1d95ccf98a5b856ec618a069f163263adf5f7

Observation 28bc5f7d-8fc1-495e-87c5-bde3063a1d2b · outbound

This paper cites Kennedy and R.

EvolKV: Evolutionary KV Cache Compression for LLM Inference Kennedy and R

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-04T20:53:29.384326Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T20:53:29.384326Z digest=sha256:a3f6229c47e064ab81089951caf323a4762cbd16b58c8608414951bff54adfc9

Observation 217b08b5-b4f7-47c9-acf3-a86a751fc34e · outbound

This paper cites an unresolved cited work.

EvolKV: Evolutionary KV Cache Compression for LLM Inference Unresolved cited work

Reference 31

Resolution
unresolved
raw_fallback, observed 2026-08-04T20:53:32.773075Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-04T20:53:29.443038Z digest=sha256:b316557d8658c2b92eaeb480ee42c933133f2deabec1db0d25d56206caa04c6a

Observation 710b9a84-f8e7-4dee-b917-7c5ebf46d614 · outbound

This paper cites The NarrativeQA Reading Comprehension Challenge.

EvolKV: Evolutionary KV Cache Compression for LLM Inference The NarrativeQA Reading Comprehension Challenge

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-04T20:53:29.536470Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T20:53:29.536470Z digest=sha256:61032303656bff72734314bf5b4594a6fcaa1d1cf4011f992fa6974e16923410

Observation 007b7138-efe7-469d-a343-b3a8c91df1dc · outbound

This paper cites an unresolved cited work.

EvolKV: Evolutionary KV Cache Compression for LLM Inference Unresolved cited work

Reference 33

Resolution
unresolved
raw_fallback, observed 2026-08-04T20:53:32.569127Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-04T20:53:29.595373Z digest=sha256:1c43c50c8d3e795b566c58418b1833dc322442f544bd4dc67fd73b995ce815b2

Observation a721037a-51e1-4aea-bf3e-cc4a3655fb36 · outbound

This paper cites SnapKV: LLM Knows What You are Looking for Before Generation.

EvolKV: Evolutionary KV Cache Compression for LLM Inference SnapKV: LLM Knows What You are Looking for Before Generation

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-04T20:53:29.655289Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T20:53:29.655289Z digest=sha256:676b92d7152936f32a89533ac0bb25ad91c2943137a80662197d34043c45a114

Observation 620f96f9-9081-457b-90e4-5d9cb31461cb · outbound

This paper cites RepoBench: Benchmarking Repository-Level Code Auto-Completion Systems.

EvolKV: Evolutionary KV Cache Compression for LLM Inference RepoBench: Benchmarking Repository-Level Code Auto-Completion Systems

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-04T20:53:29.703129Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T20:53:29.703129Z digest=sha256:e42b760a0d9f3e59650fce9c798446c757856d6db483011afc85bc06207cb60e

Observation 11cc42c4-5729-4e86-9a55-7f13fbb5c03b · outbound

This paper cites Scissorhands: Exploiting the Persistence of Importance Hypothesis for LLM KV Cache Compression at Test Time.

EvolKV: Evolutionary KV Cache Compression for LLM Inference Scissorhands: Exploiting the Persistence of Importance Hypothesis for LLM KV Cache Compression at Test Time

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-04T20:53:29.751152Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T20:53:29.751152Z digest=sha256:37bcd21745dcc452cafa39d2de1d2293ac0e1f8166ca4a7aa320d16e3e8ff77a

Observation 85e4721f-70b0-40ec-9c85-57f601f6a962 · outbound

This paper cites StarCoder 2 and The Stack v2: The Next Generation.

EvolKV: Evolutionary KV Cache Compression for LLM Inference StarCoder 2 and The Stack v2: The Next Generation

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-04T20:53:29.804774Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T20:53:29.804774Z digest=sha256:5866b6156ceb3979b22428d7860ca97a995c489bc9f3512fdab25de5d6f2c0d5

Observation 3881c240-8c37-4ffd-86b0-e6b050091b4c · outbound

This paper cites GPT-4 Technical Report.

EvolKV: Evolutionary KV Cache Compression for LLM Inference GPT-4 Technical Report

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-04T20:53:29.842871Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T20:53:29.842871Z digest=sha256:ce35c189eef29989dd66a4d305c19551b13494da9695ea49cf2711ac3992dfac

Observation f833dff9-0183-4bd6-ac51-7b99147bfd01 · outbound

This paper cites an unresolved cited work.

EvolKV: Evolutionary KV Cache Compression for LLM Inference Unresolved cited work

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-04T20:53:29.906574Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T20:53:29.906574Z digest=sha256:d20541c0f669ce618c21f1d4dbafcd066ad113b78ff0e039fced1f7f9b9728b8

Observation a2ab1ab1-e008-4d84-959a-dbcb3bf958eb · outbound

This paper cites Language Modelling with Pixels.

EvolKV: Evolutionary KV Cache Compression for LLM Inference Language Modelling with Pixels

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-04T20:53:29.956650Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T20:53:29.956650Z digest=sha256:fc6d6378143217f1f740f8f867bb992f1b9600c8934683d3da89a17fc29c1ac9

Observation 411d3c69-623e-4cf4-a814-2ca78ee496be · outbound

This paper cites Fu, Zhiqiang Xie, Beidi Chen, Clark W.

EvolKV: Evolutionary KV Cache Compression for LLM Inference Fu, Zhiqiang Xie, Beidi Chen, Clark W

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-04T20:53:32.356461Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-04T20:53:30.052751Z digest=sha256:219afe84161aec51ab7a37ff577bb1504e9d8c665de09e456b89f72fc0d5530f

Observation 64453fe3-4c28-4191-8b70-bc5a7136d4cb · outbound

This paper cites Keep the Cost Down: A Review on Methods to Optimize LLM' s KV-Cache Consumption.

EvolKV: Evolutionary KV Cache Compression for LLM Inference Keep the Cost Down: A Review on Methods to Optimize LLM' s KV-Cache Consumption

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-04T20:53:30.115688Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T20:53:30.115688Z digest=sha256:4d2ed8f25ce69155bc20682dfe3e5f31ff36ad8682a434f0e1edf6f63f36486d

Observation d8738f86-aa29-4df8-809c-ad338e73fd5c · outbound

This paper cites Layer by Layer: Uncovering Hidden Representations in Language Models.

EvolKV: Evolutionary KV Cache Compression for LLM Inference Layer by Layer: Uncovering Hidden Representations in Language Models

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-04T20:53:30.156275Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T20:53:30.156275Z digest=sha256:393ac587bf564ede44c06ccbd42955302023979be0b5274564b4d00aaef7f9ad

Observation 4c26b3d9-2503-4607-9a8c-225d265e964c · outbound

This paper cites an unresolved cited work.

EvolKV: Evolutionary KV Cache Compression for LLM Inference Unresolved cited work

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-04T20:53:30.208315Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T20:53:30.208315Z digest=sha256:446f5510a54096c06522a860be9649248f002ac4ba4092f465f2c333c8b1dfc7

Observation 2fb67ddc-9f16-427a-8fe6-591cdc4b8232 · outbound

This paper cites Llama 2: Open Foundation and Fine-Tuned Chat Models.

EvolKV: Evolutionary KV Cache Compression for LLM Inference Llama 2: Open Foundation and Fine-Tuned Chat Models

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-04T20:53:30.257649Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T20:53:30.257649Z digest=sha256:f51d83e2b5544cc6744e3ab7f4a96ac001a567390031f5431d58aa7b43ea4231

Observation 8264b62f-82ed-4454-8939-f2b0d50d5de0 · outbound

This paper cites MuSiQue: Multihop Questions via Single-hop Question Composition.

EvolKV: Evolutionary KV Cache Compression for LLM Inference MuSiQue: Multihop Questions via Single-hop Question Composition

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-04T20:53:30.299251Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T20:53:30.299251Z digest=sha256:1d2b89edaa3ff659eef9d9fc82fa383f4ae158b4ad9322688b3895c142f7eaec

Observation 514806e7-e6b7-442c-8cc4-2365e51547d8 · outbound

This paper cites an unresolved cited work.

EvolKV: Evolutionary KV Cache Compression for LLM Inference Unresolved cited work

Reference 47

Resolution
verified exact
doi, observed 2026-08-04T20:53:31.024481Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-04T20:53:30.361487Z digest=sha256:07e152968af27e7b74ee3fc8d41cb8fe5e2ec9f63a2794d4655995430f0f3a22

Observation 118dd673-bb33-4469-9a45-f1d5d9bf49c2 · outbound

This paper cites Rethinking the Value of Transformer Components.

EvolKV: Evolutionary KV Cache Compression for LLM Inference Rethinking the Value of Transformer Components

Reference 48

Resolution
verified exact
local_arxiv, observed 2026-08-04T20:53:31.795752Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-04T20:53:30.417248Z digest=sha256:e685b9c510066216d6efe35d9ddd04bcf0e63f908d8d44303a3ea439a83f7dca

Observation e9f72f18-0519-4052-abec-3a59ada6d253 · outbound

This paper cites Efficient Streaming Language Models with Attention Sinks.

EvolKV: Evolutionary KV Cache Compression for LLM Inference Efficient Streaming Language Models with Attention Sinks

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-04T20:53:30.488876Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T20:53:30.488876Z digest=sha256:f63a9c778a92a3d25ed9e3117d6f474b29f20c4f75fc29b23861d3062f23f8ed

Observation 9ac1f9ac-caa9-43a1-85fd-33854434a692 · outbound

This paper cites PyramidInfer: Pyramid KV Cache Compression for High-throughput LLM Inference.

EvolKV: Evolutionary KV Cache Compression for LLM Inference PyramidInfer: Pyramid KV Cache Compression for High-throughput LLM Inference

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-04T20:53:30.523155Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T20:53:30.523155Z digest=sha256:cc6287b7501c1d3985ebf905d5ee3e973866f5689c51f30a26644ee2a5b71cb2

Observation 03a53764-862b-4270-90a3-bc611a1c7698 · outbound

This paper cites HotpotQA: A Dataset for Diverse, Explainable Multi-hop Question Answering.

EvolKV: Evolutionary KV Cache Compression for LLM Inference HotpotQA: A Dataset for Diverse, Explainable Multi-hop Question Answering

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-04T20:53:30.564217Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T20:53:30.564217Z digest=sha256:c8212413ecef7f26284c181c84f3924fcde85c92dabdf9ba52e76f9f4e8f1af3

Observation acec151c-263f-4e90-81a4-32562f1924b2 · outbound

This paper cites Investigating Layer Importance in Large Language Models.

EvolKV: Evolutionary KV Cache Compression for LLM Inference Investigating Layer Importance in Large Language Models

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-04T20:53:30.630054Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T20:53:30.630054Z digest=sha256:fd47a9c4196fda82390a633c4e2661d31dda1a86c2ade9ece8bf2c614aa62134

Observation 451a3792-123c-43ed-b4ce-a1a3e8bebde6 · outbound

This paper cites H$_2$O: Heavy-Hitter Oracle for Efficient Generative Inference of Large Language Models.

EvolKV: Evolutionary KV Cache Compression for LLM Inference H$_2$O: Heavy-Hitter Oracle for Efficient Generative Inference of Large Language Models

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-04T20:53:30.705096Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T20:53:30.705096Z digest=sha256:6ac8ac85cb23dff150ec89fae4edd4aaab07debeac37228205d6f88541c73af5

Observation e4422975-3d6f-4edc-a58b-11aa2d9e8ac6 · outbound

This paper cites QMSum: A New Benchmark for Query-based Multi-domain Meeting Summarization.

EvolKV: Evolutionary KV Cache Compression for LLM Inference QMSum: A New Benchmark for Query-based Multi-domain Meeting Summarization

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-04T20:53:30.792140Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T20:53:30.792140Z digest=sha256:342a08c59a1d4735c122ca816fc9a694a0418e18615c7d583eb1c884eba7b25d

Pith citing papers

No inbound Pith citation observations are available.