Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 22 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 33 inbound Pith citation observations for arXiv:2410.21465.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-21T06:32:19.484+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-16T10:28:52.508979Z
A source-named dated measurement, never combined with another source.
Source: pith, observed 2026-07-10T01:36:44.234407Z
0 of 0 outbound references displayed
External citation measurements
No source-named external measurement is stored.
No outbound reference observations are available for this paper version.
Observation d00e5596-cb68-4190-a5e7-2648c9153629 · inbound
Attamba: Attending To Multi-Token States ShadowKV: KV Cache in Shadows for High-Throughput Long-Context LLM Inference
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 26a87e2a-968c-4121-9028-452a777bf42d · inbound
SCBench: A KV Cache-Centric Analysis of Long-Context Methods ShadowKV: KV Cache in Shadows for High-Throughput Long-Context LLM Inference
Reference 77
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8cb3a882-91f5-4131-810b-d02aac224e29 · inbound
A Survey on Large Language Model Acceleration based on KV Cache Management ShadowKV: KV Cache in Shadows for High-Throughput Long-Context LLM Inference
Reference 71
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a12d4335-85ec-4560-8a79-12532bb34198 · inbound
PolarQuant: Quantizing KV Caches with Polar Transformation ShadowKV: KV Cache in Shadows for High-Throughput Long-Context LLM Inference
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 858ec0b8-1b25-42de-a93e-597ba1aa812a · inbound
Speculative Prefill: Turbocharging TTFT with Lightweight and Training-Free Token Importance Estimation ShadowKV: KV Cache in Shadows for High-Throughput Long-Context LLM Inference
Reference 44
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e9cacb04-604f-45c6-97e3-80b3e285eaeb · inbound
Unveiling Simplicities of Attention: Adaptive Long-Context Head Identification ShadowKV: KV Cache in Shadows for High-Throughput Long-Context LLM Inference
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d19f214f-d719-492e-83ff-27baf0ba7ae9 · inbound
EcoServe: Enabling Cost-effective LLM Serving with Proactive Intra- and Inter-Instance Orchestration ShadowKV: KV Cache in Shadows for High-Throughput Long-Context LLM Inference
Reference 45
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2393bfa6-f739-4a74-81c7-0b8bee2b5202 · inbound
Hardware-Efficient Attention for Fast Decoding ShadowKV: KV Cache in Shadows for High-Throughput Long-Context LLM Inference
Reference 64
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2178ee9d-6c7e-4b91-bb98-be1e9a5b5e6b · inbound
HATA: Trainable and Hardware-Efficient Hash-Aware Top-k Attention for Scalable Large Model Inference ShadowKV: KV Cache in Shadows for High-Throughput Long-Context LLM Inference
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2122038e-f8b7-4b8f-8ae8-c0b6172ef580 · inbound
Kinetics: Rethinking Test-Time Scaling Laws ShadowKV: KV Cache in Shadows for High-Throughput Long-Context LLM Inference
Reference 51
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9545cbf4-34d8-46f4-9560-9cacafa53065 · inbound
Learn from the Past: Fast Sparse Indexing for Large Language Model Decoding ShadowKV: KV Cache in Shadows for High-Throughput Long-Context LLM Inference
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 30d37a71-f428-4826-b899-af7e042db1c9 · inbound
OjaKV: Context-Aware Online Low-Rank KV Cache Compression ShadowKV: KV Cache in Shadows for High-Throughput Long-Context LLM Inference
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 5ed6dcbf-1129-423e-a47c-1dd72e511631 · inbound
HeteroCache: A Dynamic Retrieval Approach to Heterogeneous KV Cache Compression for Long-Context LLM Inference ShadowKV: KV Cache in Shadows for High-Throughput Long-Context LLM Inference
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 2c23b3d6-85bb-4a6a-a169-544ac13a77e2 · inbound
ParisKV: Fast and Drift-Robust KV-Cache Retrieval for Long-Context LLMs ShadowKV: KV Cache in Shadows for High-Throughput Long-Context LLM Inference
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7e99a5d4-efe5-4e04-8563-65f6fea7ffa4 · inbound
ThunderAgent: A Simple, Fast and Program-Aware Agentic Inference System ShadowKV: KV Cache in Shadows for High-Throughput Long-Context LLM Inference
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6c2560c7-d8c1-4487-9f62-eb0da98abf70 · inbound
PrefixWall: Mitigating Prefix Caching Side Channels in Shared LLM Systems ShadowKV: KV Cache in Shadows for High-Throughput Long-Context LLM Inference
Reference 61
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation b68da320-0dbd-4fb0-a708-e93f8985d0d7 · inbound
POINTS-Long: Adaptive Dual-Mode Visual Reasoning in MLLMs ShadowKV: KV Cache in Shadows for High-Throughput Long-Context LLM Inference
Reference 76
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 264a3734-c713-461d-9c20-76f23d6b4347 · inbound
Co-evolving Agent Architectures and Interpretable Reasoning for Automated Optimization ShadowKV: KV Cache in Shadows for High-Throughput Long-Context LLM Inference
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation ef1bd31d-b45d-4b87-82fd-7d12adf05460 · inbound
DeInfer: Efficient Parallel Inferencing for Decomposed Large Language Models ShadowKV: KV Cache in Shadows for High-Throughput Long-Context LLM Inference
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation e8b5f251-9dad-492a-9869-7db3bf39207a · inbound
An Efficient Hybrid Sparse Attention with CPU-GPU Parallelism for Long-Context Inference ShadowKV: KV Cache in Shadows for High-Throughput Long-Context LLM Inference
Reference 42
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 56215eab-bc41-47e8-9332-38d08d6b5f28 · inbound
Make Each Token Count: Towards Improving Long-Context Performance with KV Cache Eviction ShadowKV: KV Cache in Shadows for High-Throughput Long-Context LLM Inference
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 2de173d6-f8da-40a1-b937-7a2692db3acb · inbound
DashAttention: Differentiable and Adaptive Sparse Hierarchical Attention ShadowKV: KV Cache in Shadows for High-Throughput Long-Context LLM Inference
Reference 52
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 063962f1-9fef-4255-8ef2-5789fd2de9ef · inbound
Idleness is Relative: Exploiting Tool-Call Idle Windows for Offloading in Agentic Systems with MORI ShadowKV: KV Cache in Shadows for High-Throughput Long-Context LLM Inference
Reference 54
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation fc0b3de4-46fa-42f1-a6d9-c72042f05787 · inbound
Dynamic Short Convolutions Improve Transformers ShadowKV: KV Cache in Shadows for High-Throughput Long-Context LLM Inference
Reference 103
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 8373b765-eb4e-468d-851a-30e80082cc0c · inbound
Vortex: Efficient and Programmable Sparse Attention Serving for AI Agents ShadowKV: KV Cache in Shadows for High-Throughput Long-Context LLM Inference
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation fa206f97-18aa-4903-b6a7-7d047684e8c8 · inbound
Predict, Reuse, and Repair: Accelerating Dynamic Sparse Attention for Long-Context LLM Decoding ShadowKV: KV Cache in Shadows for High-Throughput Long-Context LLM Inference
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 25da11a0-f559-4f99-b474-27a1f651eaf0 · inbound
From Tensor Buffer to Distributed Memory Hierarchy: A Survey of KV Cache Management for LLM Serving ShadowKV: KV Cache in Shadows for High-Throughput Long-Context LLM Inference
Reference 111
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d9e434ec-3b15-4c4e-9c3a-aad7ad305b00 · inbound
What to Keep, What to Forget: A Rate--Distortion View of Memory Compaction in LLMs and Agents ShadowKV: KV Cache in Shadows for High-Throughput Long-Context LLM Inference
Reference 109
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 27517b0b-e45e-47a2-b685-9bab0a290d64 · inbound
PhyCheck: Fine-Grained Evidence-Grounded Dataset for Physical Law Understanding in Video-LLMs ShadowKV: KV Cache in Shadows for High-Throughput Long-Context LLM Inference
Reference 75
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8a0700cc-3573-43c3-8a64-b4cf4e03cc54 · inbound
PhyCheck: Fine-Grained Evidence-Grounded Dataset for Physical Law Understanding in Video-LLMs ShadowKV: KV Cache in Shadows for High-Throughput Long-Context LLM Inference
Reference 75
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a205d2e9-b4d4-41cd-b526-f9ceeb3fcd79 · inbound
SAKI: Score-Aware Low-Rank Key Indexing with Random-Matrix Noise Correction for KV Retrieval ShadowKV: KV Cache in Shadows for High-Throughput Long-Context LLM Inference
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9437c568-adbd-4aae-a0e7-bfa78d3487cb · inbound
SAKI: Score-Aware Low-Rank Key Indexing with Random-Matrix Noise Correction for KV Retrieval ShadowKV: KV Cache in Shadows for High-Throughput Long-Context LLM Inference
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3a033333-3dc6-44ff-94f6-e1e6a600920c · inbound
OasisKV: Scaling In-Decode KV Cache Beyond HBM with Lookahead Sparse Prefetching ShadowKV: KV Cache in Shadows for High-Throughput Long-Context LLM Inference
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.