Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-05-12T05:02:25.513351Z
Paper Citation Record · LEDGER
As of 3 August 2026, this Paper Citation Record lists 38 of 38 outbound references and 1 inbound Pith citation observation for arXiv:2605.09649.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-05-12T05:02:25.513351Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-03T06:30:56.289259+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-01T02:23:36.089893Z
A source-named dated measurement, never combined with another source.
Source: cited_works
38 of 38 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 013b6995-1260-4408-ae8b-3a7b2f547ad1 · outbound
Make Each Token Count: Towards Improving Long-Context Performance with KV Cache Eviction LongBench v2: Towards Deeper Understanding and Reasoning on Realistic Long-context Multitasks
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation f56a4737-ccb5-4148-8a48-05c14368f787 · outbound
Make Each Token Count: Towards Improving Long-Context Performance with KV Cache Eviction arXiv preprint arXiv:2512.03324 , year=
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 416e77be-388e-4f64-a49b-e09f86040c65 · outbound
Make Each Token Count: Towards Improving Long-Context Performance with KV Cache Eviction PyramidKV: Dynamic KV Cache Compression based on Pyramidal Information Funneling
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 38f4cab9-de76-4f6c-85c8-ad27a820454d · outbound
Make Each Token Count: Towards Improving Long-Context Performance with KV Cache Eviction R-KV: Redundancy-aware KV cache compression for reasoning models
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation bd2546f2-e360-45c6-b995-80c4c599d8bf · outbound
Make Each Token Count: Towards Improving Long-Context Performance with KV Cache Eviction Training Verifiers to Solve Math Word Problems
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation f6d53288-ed26-40d2-aec1-2619bf91110f · outbound
Make Each Token Count: Towards Improving Long-Context Performance with KV Cache Eviction Gemini 2.5: Pushing the Frontier with Advanced Reasoning, Multimodality, Long Context, and Next Generation Agentic Capabilities
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 5ccde352-3862-424b-949e-1f850ec11e1f · outbound
Make Each Token Count: Towards Improving Long-Context Performance with KV Cache Eviction FlashAttention-2: Faster Attention with Better Parallelism and Work Partitioning
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 7bd1e6bb-a2e3-4e1f-aade-e22b7ce1ed6c · outbound
Make Each Token Count: Towards Improving Long-Context Performance with KV Cache Eviction How Sparse Attention Approximates Exact Attention? Your Attention is Naturally $n^C$-Sparse
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 54352432-fa9f-43b7-a2ed-e73125c78673 · outbound
Make Each Token Count: Towards Improving Long-Context Performance with KV Cache Eviction Ada-KV: Optimizing KV Cache Eviction by Adaptive Budget Allocation for Efficient LLM Inference
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 4e375f8d-a619-4f26-8b6a-d8c8893f666e · outbound
Make Each Token Count: Towards Improving Long-Context Performance with KV Cache Eviction SeerAttention-R: Sparse Attention Adaptation for Long Reasoning
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation b238eeaf-a668-4850-a0a1-a91ca84d10dc · outbound
Make Each Token Count: Towards Improving Long-Context Performance with KV Cache Eviction Dialogue Without Limits: Constant-Sized KV Caches for Extended Responses in LLMs
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 5b879935-f2aa-4a79-a91e-c62fa0d794b6 · outbound
Make Each Token Count: Towards Improving Long-Context Performance with KV Cache Eviction LM-Infinite: Zero-Shot Extreme Length Generalization for Large Language Models
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 786afdd9-8099-401c-a79d-856af36a9362 · outbound
Make Each Token Count: Towards Improving Long-Context Performance with KV Cache Eviction Measuring Mathematical Problem Solving With the MATH Dataset
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 3bbf2dbb-27cc-47c6-be66-704261458a7b · outbound
Make Each Token Count: Towards Improving Long-Context Performance with KV Cache Eviction Video-MMMU: Evaluating Knowledge Acquisition from Multi-Discipline Professional Videos
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 1707df9e-d9c0-4a08-b85d-db4c98a22616 · outbound
Make Each Token Count: Towards Improving Long-Context Performance with KV Cache Eviction Vision-R1: Incentivizing Reasoning Capability in Multimodal Large Language Models
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation c04ca568-05a7-4eac-b93e-69b5335d3677 · outbound
Make Each Token Count: Towards Improving Long-Context Performance with KV Cache Eviction LLaVA-NeXT-Interleave: Tackling Multi-image, Video, and 3D in Large Multimodal Models
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 85741ffd-ef11-49db-9dd8-571b56e02b91 · outbound
Make Each Token Count: Towards Improving Long-Context Performance with KV Cache Eviction RetrievalAttention: Accelerating Long-Context LLM Inference via Vector Retrieval
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 546364dd-325b-4ad1-a1f1-8c5b47a9126e · outbound
Make Each Token Count: Towards Improving Long-Context Performance with KV Cache Eviction Dynamic Memory Compression: Retrofitting LLMs for Accelerated Inference
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation e09839fd-0a65-413a-8973-1ca43760afae · outbound
Make Each Token Count: Towards Improving Long-Context Performance with KV Cache Eviction Kediff: Key similarity-based KV cache eviction for long-context LLM inference in resource-constrained environments
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 6328dd7b-6d31-48cb-9a7e-0d14b2507d5f · outbound
Make Each Token Count: Towards Improving Long-Context Performance with KV Cache Eviction Daniela Rus
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 5df17a2f-7838-4dcb-8173-216b94145f06 · outbound
Make Each Token Count: Towards Improving Long-Context Performance with KV Cache Eviction VideoMathQA: Benchmarking Mathematical Reasoning via Multimodal Understanding in Videos
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 74f78f2a-04be-427f-a45c-f635d39eb2d6 · outbound
Make Each Token Count: Towards Improving Long-Context Performance with KV Cache Eviction Vision-language- action models: Concepts, progress, applications and chal- lenges.arXiv preprint arXiv:2505.04769
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 56215eab-bc41-47e8-9332-38d08d6b5f28 · outbound
Make Each Token Count: Towards Improving Long-Context Performance with KV Cache Eviction ShadowKV: KV Cache in Shadows for High-Throughput Long-Context LLM Inference
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 037032b0-6783-450e-b744-9a1f0f70bc3a · outbound
Make Each Token Count: Towards Improving Long-Context Performance with KV Cache Eviction Quest: Query-Aware Sparsity for Efficient Long-Context LLM Inference
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation abc86ec3-430d-4170-84e6-229078555ee7 · outbound
Make Each Token Count: Towards Improving Long-Context Performance with KV Cache Eviction Gemini: A Family of Highly Capable Multimodal Models
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 2542ebda-0081-4fd9-b658-cd54ead99e4f · outbound
Make Each Token Count: Towards Improving Long-Context Performance with KV Cache Eviction VL-Cache: Sparsity and Modality-Aware KV Cache Compression for Vision-Language Model Inference Acceleration
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 8d4df5aa-921c-45b6-8a7e-2d4d6e7d8870 · outbound
Make Each Token Count: Towards Improving Long-Context Performance with KV Cache Eviction MEDA: Dynamic KV Cache Allocation for Efficient Multimodal Long-Context Inference
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation bb7f7fc5-0671-4d8e-9eff-b6e8f15902a7 · outbound
Make Each Token Count: Towards Improving Long-Context Performance with KV Cache Eviction LLMs Know What to Drop: Self-Attention Guided KV Cache Eviction for Efficient Long-Context Inference
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation d3afb04a-5cda-497d-ab69-06d54b324724 · outbound
Make Each Token Count: Towards Improving Long-Context Performance with KV Cache Eviction Stop Looking for Important Tokens in Multimodal Language Models: Duplication Matters More
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation dcc96a6a-2556-4e52-b608-a0959b909e93 · outbound
Make Each Token Count: Towards Improving Long-Context Performance with KV Cache Eviction Efficient Streaming Language Models with Attention Sinks
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation a89fbc46-e19b-4389-b915-245443a00efc · outbound
Make Each Token Count: Towards Improving Long-Context Performance with KV Cache Eviction R1-Onevision: Advancing Generalized Multimodal Reasoning through Cross-Modal Formalization
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation e3ebbf7c-c88d-474c-9976-b1d2d5ffe071 · outbound
Make Each Token Count: Towards Improving Long-Context Performance with KV Cache Eviction doi: 10.18653/v1/2025.acl-long.736
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation b425d45b-25d9-4bef-9270-e7c6144f21a2 · outbound
Make Each Token Count: Towards Improving Long-Context Performance with KV Cache Eviction LMMs-Eval: Reality Check on the Evaluation of Large Multimodal Models
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 14a8b709-b2b2-40bd-a491-717c9768bcae · outbound
Make Each Token Count: Towards Improving Long-Context Performance with KV Cache Eviction SparseVLM: Visual Token Sparsification for Efficient Vision-Language Model Inference
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 27aa9edd-9a3a-44da-b882-a9fd4188f70e · outbound
Make Each Token Count: Towards Improving Long-Context Performance with KV Cache Eviction Make Each Token Count: Towards Improving Long-Context Performance with KV Cache Eviction
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 407be0f2-ccca-4ef8-b15e-b62afb385155 · outbound
Make Each Token Count: Towards Improving Long-Context Performance with KV Cache Eviction Visual Perception
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 7820a812-dab0-40ed-acee-c8b76e865bcb · outbound
Make Each Token Count: Towards Improving Long-Context Performance with KV Cache Eviction Unresolved cited work
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation acd79c82-aa15-4f2f-9358-90700de6edb8 · outbound
Make Each Token Count: Towards Improving Long-Context Performance with KV Cache Eviction Results are averaged over 5 random seeds
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation caa0a936-0ccb-42f8-95e6-f35a53116123 · inbound
Seen, Said, or Forgotten? A Causal Audit of Visual KV Memory Across Dialog Turns Make Each Token Count: Towards Improving Long-Context Performance with KV Cache Eviction
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.