Pith. sign in

Paper Citation Record · LEDGER

ShadowKV: KV Cache in Shadows for High-Throughput Long-Context LLM Inference

As of 22 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 33 inbound Pith citation observations for arXiv:2410.21465.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2410.21465 v3

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 33 of 33 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-21T06:32:19.484+00:00

measured 33 of 33 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-16T10:28:52.508979Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-07-10T01:36:44.234407Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation d00e5596-cb68-4190-a5e7-2648c9153629 · inbound

Attamba: Attending To Multi-Token States cites this paper.

Attamba: Attending To Multi-Token States ShadowKV: KV Cache in Shadows for High-Throughput Long-Context LLM Inference

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-12T11:56:33.082127Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T11:56:33.082127Z digest=sha256:a2c8f15c1b532f66f0bc6b434b050e47bc75a2037282fece8062d2d272a31df3

Observation 26a87e2a-968c-4121-9028-452a777bf42d · inbound

SCBench: A KV Cache-Centric Analysis of Long-Context Methods cites this paper.

SCBench: A KV Cache-Centric Analysis of Long-Context Methods ShadowKV: KV Cache in Shadows for High-Throughput Long-Context LLM Inference

Reference 77

Resolution
unresolved
no resolver link, observed 2026-08-11T16:14:06.050044Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T16:14:06.050044Z digest=sha256:e637221701396651c272c20a15c8966bbfb667ff1d8114516a021a64ca174ab6

Observation 8cb3a882-91f5-4131-810b-d02aac224e29 · inbound

A Survey on Large Language Model Acceleration based on KV Cache Management cites this paper.

A Survey on Large Language Model Acceleration based on KV Cache Management ShadowKV: KV Cache in Shadows for High-Throughput Long-Context LLM Inference

Reference 71

Resolution
unresolved
no resolver link, observed 2026-08-11T00:38:46.981217Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T00:38:46.981217Z digest=sha256:20acbf334b61fe9b0cd16c40d0f8857e3f2f316547df60ca5f5cb4d051199681

Observation a12d4335-85ec-4560-8a79-12532bb34198 · inbound

PolarQuant: Quantizing KV Caches with Polar Transformation cites this paper.

PolarQuant: Quantizing KV Caches with Polar Transformation ShadowKV: KV Cache in Shadows for High-Throughput Long-Context LLM Inference

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-09T13:26:55.444082Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T13:26:55.444082Z digest=sha256:fe7b914e80cc11a19c64af687643c99d4621e5cdca9b26f6f504953590ad2ddf

Observation 858ec0b8-1b25-42de-a93e-597ba1aa812a · inbound

Speculative Prefill: Turbocharging TTFT with Lightweight and Training-Free Token Importance Estimation cites this paper.

Speculative Prefill: Turbocharging TTFT with Lightweight and Training-Free Token Importance Estimation ShadowKV: KV Cache in Shadows for High-Throughput Long-Context LLM Inference

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-09T11:11:17.779908Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T11:11:17.779908Z digest=sha256:af1280f340f025781cad5303a70fae8fb10c1f2edbe46d5691051e1e74360c91

Observation e9cacb04-604f-45c6-97e3-80b3e285eaeb · inbound

Unveiling Simplicities of Attention: Adaptive Long-Context Head Identification cites this paper.

Unveiling Simplicities of Attention: Adaptive Long-Context Head Identification ShadowKV: KV Cache in Shadows for High-Throughput Long-Context LLM Inference

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-08T13:46:11.424138Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T13:46:11.424138Z digest=sha256:e2776b60d0f614486a8d294bbe586ff5f5a222396e011648c8a15d5fffe209f2

Observation d19f214f-d719-492e-83ff-27baf0ba7ae9 · inbound

EcoServe: Enabling Cost-effective LLM Serving with Proactive Intra- and Inter-Instance Orchestration cites this paper.

EcoServe: Enabling Cost-effective LLM Serving with Proactive Intra- and Inter-Instance Orchestration ShadowKV: KV Cache in Shadows for High-Throughput Long-Context LLM Inference

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-16T10:28:52.508979Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T10:28:52.508979Z digest=sha256:7b933ab683f50c61b2ca3c504bd372f34add5ba71f8baa0104749ada79d267f4

Observation 2393bfa6-f739-4a74-81c7-0b8bee2b5202 · inbound

Hardware-Efficient Attention for Fast Decoding cites this paper.

Hardware-Efficient Attention for Fast Decoding ShadowKV: KV Cache in Shadows for High-Throughput Long-Context LLM Inference

Reference 64

Resolution
unresolved
no resolver link, observed 2026-08-07T13:32:32.885907Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:32:32.885907Z digest=sha256:9acb2fb3672e8b0c4d78bbf637ae046b2c8459a65ba48ef4ddd9746d7cbcb137

Observation 2178ee9d-6c7e-4b91-bb98-be1e9a5b5e6b · inbound

HATA: Trainable and Hardware-Efficient Hash-Aware Top-k Attention for Scalable Large Model Inference cites this paper.

HATA: Trainable and Hardware-Efficient Hash-Aware Top-k Attention for Scalable Large Model Inference ShadowKV: KV Cache in Shadows for High-Throughput Long-Context LLM Inference

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-07T11:26:00.895653Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:26:00.895653Z digest=sha256:258502a4e2e2239681870fb3fd7f0f8f7cd6907a5fb7066a9d12a998f743396b

Observation 2122038e-f8b7-4b8f-8ae8-c0b6172ef580 · inbound

Kinetics: Rethinking Test-Time Scaling Laws cites this paper.

Kinetics: Rethinking Test-Time Scaling Laws ShadowKV: KV Cache in Shadows for High-Throughput Long-Context LLM Inference

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-07T10:30:34.376950Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T10:30:34.376950Z digest=sha256:4fa1992fec9374ac0be172d60ffc5acd4aa49b997776b901bd994db349be87f4

Observation 9545cbf4-34d8-46f4-9560-9cacafa53065 · inbound

Learn from the Past: Fast Sparse Indexing for Large Language Model Decoding cites this paper.

Learn from the Past: Fast Sparse Indexing for Large Language Model Decoding ShadowKV: KV Cache in Shadows for High-Throughput Long-Context LLM Inference

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-07T12:38:39.632271Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:38:39.632271Z digest=sha256:67963f0cdaa20fccb8e773ac771a5c8a1f3867e9c30c6b91f31b6d9fa5d8a984

Observation 30d37a71-f428-4826-b899-af7e042db1c9 · inbound

OjaKV: Context-Aware Online Low-Rank KV Cache Compression cites this paper.

OjaKV: Context-Aware Online Low-Rank KV Cache Compression ShadowKV: KV Cache in Shadows for High-Throughput Long-Context LLM Inference

Reference 14

Resolution
verified exact
arxiv_id, observed 2026-05-18T13:26:24.732629Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-18T13:26:02.980973Z digest=sha256:4920f60d665fa47a1c1af0f12bd180397c47fc485de06e392ee3ff67778871a1

Observation 5ed6dcbf-1129-423e-a47c-1dd72e511631 · inbound

HeteroCache: A Dynamic Retrieval Approach to Heterogeneous KV Cache Compression for Long-Context LLM Inference cites this paper.

HeteroCache: A Dynamic Retrieval Approach to Heterogeneous KV Cache Compression for Long-Context LLM Inference ShadowKV: KV Cache in Shadows for High-Throughput Long-Context LLM Inference

Reference 9

Resolution
metadata mismatch
arxiv_id, observed 2026-05-16T13:17:54.885057Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-16T13:16:31.568604Z digest=sha256:c18a534ad7ee929c4c5bd58fd17b8536eda1fbec0eaa8cc133dacfbdf304c2da

Observation 2c23b3d6-85bb-4a6a-a169-544ac13a77e2 · inbound

ParisKV: Fast and Drift-Robust KV-Cache Retrieval for Long-Context LLMs cites this paper.

ParisKV: Fast and Drift-Robust KV-Cache Retrieval for Long-Context LLMs ShadowKV: KV Cache in Shadows for High-Throughput Long-Context LLM Inference

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-03T03:37:50.132428Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T03:37:50.132428Z digest=sha256:a1ab9198f5700f8ee35e59017a16876e764ad5f755f9165662d162b93de4a943

Observation 7e99a5d4-efe5-4e04-8563-65f6fea7ffa4 · inbound

ThunderAgent: A Simple, Fast and Program-Aware Agentic Inference System cites this paper.

ThunderAgent: A Simple, Fast and Program-Aware Agentic Inference System ShadowKV: KV Cache in Shadows for High-Throughput Long-Context LLM Inference

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-02T23:32:13.355912Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T23:32:13.355912Z digest=sha256:e55a844f54d1f828dc3b4025a9d203f6e565eec6938eb19c9011969d9a2181f7

Observation 6c2560c7-d8c1-4487-9f62-eb0da98abf70 · inbound

PrefixWall: Mitigating Prefix Caching Side Channels in Shared LLM Systems cites this paper.

PrefixWall: Mitigating Prefix Caching Side Channels in Shared LLM Systems ShadowKV: KV Cache in Shadows for High-Throughput Long-Context LLM Inference

Reference 61

Resolution
verified exact
arxiv_id, observed 2026-05-21T12:15:06.756004Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-21T12:14:09.509302Z digest=sha256:46c5a872a0dde3edf7a0ef86be84ff3f44adb6fb90ecd5b704ad8a9f06c8395c

Observation b68da320-0dbd-4fb0-a708-e93f8985d0d7 · inbound

POINTS-Long: Adaptive Dual-Mode Visual Reasoning in MLLMs cites this paper.

POINTS-Long: Adaptive Dual-Mode Visual Reasoning in MLLMs ShadowKV: KV Cache in Shadows for High-Throughput Long-Context LLM Inference

Reference 76

Resolution
verified exact
arxiv_id, observed 2026-05-11T10:41:04.055209Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-10T15:23:08.671342Z digest=sha256:4d22a20b57a5c1e64dcbe3d2b376f7e09e846f80df529b8f3973ae556a16f714

Observation 264a3734-c713-461d-9c20-76f23d6b4347 · inbound

Co-evolving Agent Architectures and Interpretable Reasoning for Automated Optimization cites this paper.

Co-evolving Agent Architectures and Interpretable Reasoning for Automated Optimization ShadowKV: KV Cache in Shadows for High-Throughput Long-Context LLM Inference

Reference 16

Resolution
verified exact
local_arxiv, observed 2026-07-05T16:31:16.090941Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-07-05T16:27:03.569327Z digest=sha256:47178224e14faf7afa17fa8934f517492ede25a8f4d84471e3af4e2f235fe132

Observation ef1bd31d-b45d-4b87-82fd-7d12adf05460 · inbound

DeInfer: Efficient Parallel Inferencing for Decomposed Large Language Models cites this paper.

DeInfer: Efficient Parallel Inferencing for Decomposed Large Language Models ShadowKV: KV Cache in Shadows for High-Throughput Long-Context LLM Inference

Reference 16

Resolution
verified exact
arxiv_id, observed 2026-05-10T05:41:02.125976Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-10T05:38:17.542951Z digest=sha256:681c04fa05000281d7286bde018864ee938fbc59abdeb197ad3d128a4e2dc481

Observation e8b5f251-9dad-492a-9869-7db3bf39207a · inbound

An Efficient Hybrid Sparse Attention with CPU-GPU Parallelism for Long-Context Inference cites this paper.

An Efficient Hybrid Sparse Attention with CPU-GPU Parallelism for Long-Context Inference ShadowKV: KV Cache in Shadows for High-Throughput Long-Context LLM Inference

Reference 42

Resolution
verified exact
arxiv_id, observed 2026-05-11T03:05:54.076381Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-11T02:56:28.828593Z digest=sha256:516d03c2d63f67b973c3e6d08d41bf16c2cde8e089d9dc88c5dc0566c3c2ed14

Observation 56215eab-bc41-47e8-9332-38d08d6b5f28 · inbound

Make Each Token Count: Towards Improving Long-Context Performance with KV Cache Eviction cites this paper.

Make Each Token Count: Towards Improving Long-Context Performance with KV Cache Eviction ShadowKV: KV Cache in Shadows for High-Throughput Long-Context LLM Inference

Reference 23

Resolution
verified exact
arxiv_id, observed 2026-05-12T05:41:26.478146Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-12T05:02:25.513351Z digest=sha256:dd4d5bf49e608eb546783adfee6f9c02f96e7677971656d18a5ae8e19a5fbc6e

Observation 2de173d6-f8da-40a1-b937-7a2692db3acb · inbound

DashAttention: Differentiable and Adaptive Sparse Hierarchical Attention cites this paper.

DashAttention: Differentiable and Adaptive Sparse Hierarchical Attention ShadowKV: KV Cache in Shadows for High-Throughput Long-Context LLM Inference

Reference 52

Resolution
verified exact
arxiv_id, observed 2026-05-20T10:53:13.593270Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-20T10:50:12.926232Z digest=sha256:a6aa00446dcef283fdd3a6779dd4e38f5a77ae3f5556ec062e8ee8d8c32d678b

Observation 063962f1-9fef-4255-8ef2-5789fd2de9ef · inbound

Idleness is Relative: Exploiting Tool-Call Idle Windows for Offloading in Agentic Systems with MORI cites this paper.

Idleness is Relative: Exploiting Tool-Call Idle Windows for Offloading in Agentic Systems with MORI ShadowKV: KV Cache in Shadows for High-Throughput Long-Context LLM Inference

Reference 54

Resolution
verified exact
arxiv_id, observed 2026-07-01T21:06:13.602579Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-06-28T17:30:56.324289Z digest=sha256:917a7e118183c7026d4bc5569fae8c6c756df2abf78478c956d8af52052f2188

Observation fc0b3de4-46fa-42f1-a6d9-c72042f05787 · inbound

Dynamic Short Convolutions Improve Transformers cites this paper.

Dynamic Short Convolutions Improve Transformers ShadowKV: KV Cache in Shadows for High-Throughput Long-Context LLM Inference

Reference 103

Resolution
metadata mismatch
arxiv_id, observed 2026-07-02T02:36:26.995018Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-06-28T10:48:50.103004Z digest=sha256:6b78acc97d6f59e1e6ff094e7c3386638d5350fa8945b662db62def904613126

Observation 8373b765-eb4e-468d-851a-30e80082cc0c · inbound

Vortex: Efficient and Programmable Sparse Attention Serving for AI Agents cites this paper.

Vortex: Efficient and Programmable Sparse Attention Serving for AI Agents ShadowKV: KV Cache in Shadows for High-Throughput Long-Context LLM Inference

Reference 32

Resolution
verified exact
arxiv_id, observed 2026-07-02T13:36:59.437987Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-06-28T01:07:14.691347Z digest=sha256:5c9a864af1340d11049cb04a6ca2cd6e4d4405d0a3f1cf54183232dfae9897d3

Observation fa206f97-18aa-4903-b6a7-7d047684e8c8 · inbound

Predict, Reuse, and Repair: Accelerating Dynamic Sparse Attention for Long-Context LLM Decoding cites this paper.

Predict, Reuse, and Repair: Accelerating Dynamic Sparse Attention for Long-Context LLM Decoding ShadowKV: KV Cache in Shadows for High-Throughput Long-Context LLM Inference

Reference 36

Resolution
metadata mismatch
arxiv_id, observed 2026-06-30T07:04:20.907471Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-06-30T07:03:08.617257Z digest=sha256:54dd328acb8af58b61283c36273590e104ab000b345008c20b5c597563a81cb3

Observation 25da11a0-f559-4f99-b474-27a1f651eaf0 · inbound

From Tensor Buffer to Distributed Memory Hierarchy: A Survey of KV Cache Management for LLM Serving cites this paper.

From Tensor Buffer to Distributed Memory Hierarchy: A Survey of KV Cache Management for LLM Serving ShadowKV: KV Cache in Shadows for High-Throughput Long-Context LLM Inference

Reference 111

Resolution
unresolved
no resolver link, observed 2026-07-12T09:50:23.266920Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T09:50:23.266920Z digest=sha256:d7a354857919ad3d6a2f9a6451bb3e991f17ddb545a0c43341842fa441b28e5a

Observation d9e434ec-3b15-4c4e-9c3a-aad7ad305b00 · inbound

What to Keep, What to Forget: A Rate--Distortion View of Memory Compaction in LLMs and Agents cites this paper.

What to Keep, What to Forget: A Rate--Distortion View of Memory Compaction in LLMs and Agents ShadowKV: KV Cache in Shadows for High-Throughput Long-Context LLM Inference

Reference 109

Resolution
verified exact
local_arxiv, observed 2026-07-10T01:36:44.235599Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-07-10T01:26:59.421158Z digest=sha256:37dc446692a364715ad2b667bba71e186c347ef6e79009ab2d68232334f2fc59

Observation 27517b0b-e45e-47a2-b685-9bab0a290d64 · inbound

PhyCheck: Fine-Grained Evidence-Grounded Dataset for Physical Law Understanding in Video-LLMs cites this paper.

PhyCheck: Fine-Grained Evidence-Grounded Dataset for Physical Law Understanding in Video-LLMs ShadowKV: KV Cache in Shadows for High-Throughput Long-Context LLM Inference

Reference 75

Resolution
unresolved
no resolver link, observed 2026-08-04T13:43:55.710075Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T13:43:55.710075Z digest=sha256:626010651e27cc59455e7c6844fc5bb8207d4afa95a1080588fb4cca7ba9070e

Observation 8a0700cc-3573-43c3-8a64-b4cf4e03cc54 · inbound

PhyCheck: Fine-Grained Evidence-Grounded Dataset for Physical Law Understanding in Video-LLMs cites this paper.

PhyCheck: Fine-Grained Evidence-Grounded Dataset for Physical Law Understanding in Video-LLMs ShadowKV: KV Cache in Shadows for High-Throughput Long-Context LLM Inference

Reference 75

Resolution
unresolved
no resolver link, observed 2026-08-07T00:11:49.064199Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:11:49.064199Z digest=sha256:83b7383d6717bcc823648e164dc4e69b502812fff6dd9c3ebb605fad97a46986

Observation a205d2e9-b4d4-41cd-b526-f9ceeb3fcd79 · inbound

SAKI: Score-Aware Low-Rank Key Indexing with Random-Matrix Noise Correction for KV Retrieval cites this paper.

SAKI: Score-Aware Low-Rank Key Indexing with Random-Matrix Noise Correction for KV Retrieval ShadowKV: KV Cache in Shadows for High-Throughput Long-Context LLM Inference

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-05T23:10:35.446082Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T23:10:35.446082Z digest=sha256:3acc2368b59e2d46a0dad265cd437d13f7f4d342f4d0c4a1811cb94b22bec8e3

Observation 9437c568-adbd-4aae-a0e7-bfa78d3487cb · inbound

SAKI: Score-Aware Low-Rank Key Indexing with Random-Matrix Noise Correction for KV Retrieval cites this paper.

SAKI: Score-Aware Low-Rank Key Indexing with Random-Matrix Noise Correction for KV Retrieval ShadowKV: KV Cache in Shadows for High-Throughput Long-Context LLM Inference

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-08T00:54:42.703743Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T00:54:42.703743Z digest=sha256:b4e676872c476e9a9f1fadf5aab10b3ddea549cc90bd9210b6ca7836e8a612d6

Observation 3a033333-3dc6-44ff-94f6-e1e6a600920c · inbound

OasisKV: Scaling In-Decode KV Cache Beyond HBM with Lookahead Sparse Prefetching cites this paper.

OasisKV: Scaling In-Decode KV Cache Beyond HBM with Lookahead Sparse Prefetching ShadowKV: KV Cache in Shadows for High-Throughput Long-Context LLM Inference

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-12T00:32:37.689283Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T00:32:37.689283Z digest=sha256:aebc6bdaf1c3ee3c137f2fff9496d2b443da4868f2b05a510838ed244c348f5f