Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-07-14T10:41:08.967261Z
Paper Citation Record · LEDGER
As of 21 August 2026, this Paper Citation Record lists 33 of 33 outbound references and 1 inbound Pith citation observation for arXiv:2607.10582.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-07-14T10:41:08.967261Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-21T06:32:19.484+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-12T00:50:51.344226Z
A source-named dated measurement, never combined with another source.
Source: pith, observed 2026-08-12T00:50:51.451312Z
33 of 33 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 4f36011c-a7ce-48f4-803d-a4eea53f44fe · outbound
MemDecay: Region-Aware KV Cache Eviction for Efficient LLM Agent Inference Efficient memory management for large language model serving with PagedAttention,
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5c92b6ea-b32b-4744-8896-2ca543c1da39 · outbound
MemDecay: Region-Aware KV Cache Eviction for Efficient LLM Agent Inference A Survey on Large Language Model Acceleration based on KV Cache Management
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 350e4f42-fdfe-4fe7-9b88-21ae88e08b04 · outbound
MemDecay: Region-Aware KV Cache Eviction for Efficient LLM Agent Inference Model tells you what to discard: Adaptive KV cache compression for LLMs,
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fd1634c4-70f1-4929-ba94-e707d266f298 · outbound
MemDecay: Region-Aware KV Cache Eviction for Efficient LLM Agent Inference Ada-KV: Optimizing KV Cache Eviction by Adaptive Budget Allocation for Efficient LLM Inference
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9afaa519-d911-498c-8373-6776ba34ea27 · outbound
MemDecay: Region-Aware KV Cache Eviction for Efficient LLM Agent Inference H 2O: Heavy-hitter oracle for efficient generative inference of large language models,
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0434c276-b9ab-4dcf-bb36-d9d944ae0d46 · outbound
MemDecay: Region-Aware KV Cache Eviction for Efficient LLM Agent Inference Efficient streaming language models with attention sinks,
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ea27f054-3853-440c-a141-84780c5d90fd · outbound
MemDecay: Region-Aware KV Cache Eviction for Efficient LLM Agent Inference Transformers are Multi-State RNNs
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b25f484d-0b96-46b8-b283-fe320d042e79 · outbound
MemDecay: Region-Aware KV Cache Eviction for Efficient LLM Agent Inference A Survey on the Memory Mechanism of Large Language Model based Agents
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cf72e7fd-b383-499b-a431-a20a9ebec4c2 · outbound
MemDecay: Region-Aware KV Cache Eviction for Efficient LLM Agent Inference Not All Tokens Are Worth Caching: Learning Semantic-Aware Eviction for LLM Prefix Caches
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b0f1d5fc-036e-4490-94a9-110c32705c81 · outbound
MemDecay: Region-Aware KV Cache Eviction for Efficient LLM Agent Inference RoleKV: Role-aware KV cache management for the inverted age-importance of LLM agent context,
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 08a7961a-011b-4b6f-b790-034d015d43d7 · outbound
MemDecay: Region-Aware KV Cache Eviction for Efficient LLM Agent Inference ChunkAttention: Efficient Self-Attention with Prefix-Aware KV Cache and Two-Phase Partition
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e0f38406-fbfb-4440-bd3d-c2159b5ba982 · outbound
MemDecay: Region-Aware KV Cache Eviction for Efficient LLM Agent Inference CacheBlend: Fast large language model serving for RAG with cached knowledge fusion,
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0e4738c2-19df-4315-b1db-5f462903a60d · outbound
MemDecay: Region-Aware KV Cache Eviction for Efficient LLM Agent Inference KV Cache Compression for Inference Efficiency in LLMs: A Review
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d8546e18-5f2d-4ccc-8860-07e5dbe81e06 · outbound
MemDecay: Region-Aware KV Cache Eviction for Efficient LLM Agent Inference Scissorhands: Exploiting the persistence of importance hypothesis for LLM KV cache compression at test time,
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0340ab23-ea33-4827-9a37-e72c97a19b05 · outbound
MemDecay: Region-Aware KV Cache Eviction for Efficient LLM Agent Inference SnapKV: LLM Knows What You are Looking for Before Generation
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f582c2a5-29ed-435d-ac7a-56ce24de0cb0 · outbound
MemDecay: Region-Aware KV Cache Eviction for Efficient LLM Agent Inference NACL: A general and effective KV cache eviction framework for LLM at inference time,
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 79c7c93a-1b6c-4c10-b99a-32b14650c623 · outbound
MemDecay: Region-Aware KV Cache Eviction for Efficient LLM Agent Inference Expected attention: KV cache compression by estimating attention from future queries distribution,
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c0ad32f2-a22d-4f65-893a-82da3d67144b · outbound
MemDecay: Region-Aware KV Cache Eviction for Efficient LLM Agent Inference Quest: Query-aware sparsity for efficient long-context LLM inference,
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 899d79c0-e253-414d-abfe-02877ba9f3c5 · outbound
MemDecay: Region-Aware KV Cache Eviction for Efficient LLM Agent Inference ClusterKV: Manipulating LLM KV Cache in Semantic Space for Recallable Compression
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 09a071a8-8537-47cb-a7d7-f81b57e9dd55 · outbound
MemDecay: Region-Aware KV Cache Eviction for Efficient LLM Agent Inference ChunkKV: Semantic-preserving KV cache compression for efficient long-context LLM inference,
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ae54892c-c64d-41f2-bc24-d762eaa8d58e · outbound
MemDecay: Region-Aware KV Cache Eviction for Efficient LLM Agent Inference KIVI: A tuning-free asymmetric 2bit quantization for KV cache,
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8ae4446e-6b58-4ce4-b870-3d85dd690c94 · outbound
MemDecay: Region-Aware KV Cache Eviction for Efficient LLM Agent Inference LKV: End-to-End Learning of Head-wise Budgets and Token Selection for LLM KV Cache Eviction
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ee8fa758-ca72-45e3-a305-147a6dec4270 · outbound
MemDecay: Region-Aware KV Cache Eviction for Efficient LLM Agent Inference Continuum: Efficient and Robust Multi-Turn LLM Agent Scheduling with KV Cache Time-to-Live
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 344ec41a-2574-4447-87ef-6a614fa5d03d · outbound
MemDecay: Region-Aware KV Cache Eviction for Efficient LLM Agent Inference KVFlow: Efficient prefix caching for accelerating LLM- based multi-agent workflows,
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1e096443-a258-4b37-ba69-98679a8bdd40 · outbound
MemDecay: Region-Aware KV Cache Eviction for Efficient LLM Agent Inference IntentKV: Cross-Turn Intent-Aware KV Cache Pruning for Agent Inference
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3dacbf59-0cdd-4c13-8f8a-ce598e94aa5e · outbound
MemDecay: Region-Aware KV Cache Eviction for Efficient LLM Agent Inference Cache what lasts: Token retention for memory-bounded KV cache in LLMs,
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 18a71781-cbfb-4810-94b8-2b628db851d9 · outbound
MemDecay: Region-Aware KV Cache Eviction for Efficient LLM Agent Inference From Human Memory to AI Memory: A Survey on Memory Mechanisms in the Era of LLMs
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9b41a0b9-8670-4a03-8d63-59bc1651e080 · outbound
MemDecay: Region-Aware KV Cache Eviction for Efficient LLM Agent Inference CONF-KV: Confidence-Aware KV Cache Eviction with Mixed-Precision Storage for Long-Horizon LLM
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a2b0cd05-6a7d-4054-aa8b-ca6dc843c250 · outbound
MemDecay: Region-Aware KV Cache Eviction for Efficient LLM Agent Inference DynamicKV: Task-Aware Adaptive KV Cache Compression for Long Context LLMs
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c49b4022-9cd6-497f-9911-c2960239c642 · outbound
MemDecay: Region-Aware KV Cache Eviction for Efficient LLM Agent Inference KVzip: Query-agnostic KV cache compression with context reconstruction,
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8ca68255-496c-411d-bebb-3f3b27b86f9b · outbound
MemDecay: Region-Aware KV Cache Eviction for Efficient LLM Agent Inference Improving WWW proxies performance with Greedy- Dual-Size-Frequency caching policy,
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 613b8f58-3c3b-4404-815b-073f1d384a6b · outbound
MemDecay: Region-Aware KV Cache Eviction for Efficient LLM Agent Inference RFC: Agent-aware KV cache phase 1 for agen- tic workloads,
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f7aa01a4-b1d6-4c36-9653-ffaa9fa45421 · outbound
MemDecay: Region-Aware KV Cache Eviction for Efficient LLM Agent Inference KV Cache Compression and Its Infra Problems,
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6f3c68d1-2b93-42c5-a5c9-389ce871a93b · inbound
CommitKV: Lifecycle-Aware KV Cache Compression via Commit Transitions for Multi-Turn Agents MemDecay: Region-Aware KV Cache Eviction for Efficient LLM Agent Inference
Reference 183
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.