Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 14 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 17 inbound Pith citation observations for arXiv:2310.07240.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-14T06:32:32.682623+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-12T17:00:04.086265Z
A source-named dated measurement, never combined with another source.
Source: pith, observed 2026-07-10T01:36:44.087990Z
0 of 0 outbound references displayed
External citation measurements
No source-named external measurement is stored.
No outbound reference observations are available for this paper version.
Observation 78ac79e8-f742-4930-865d-5fec79050c14 · inbound
LLMSteer: Improving Long-Context LLM Inference by Steering Attention on Reused Contexts CacheGen: KV Cache Compression and Streaming for Fast Large Language Model Serving
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation febecb8e-2517-4865-896c-d5737ddc14a5 · inbound
KVDirect: Distributed Disaggregated LLM Inference CacheGen: KV Cache Compression and Streaming for Fast Large Language Model Serving
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 12d1fef0-7650-4007-95a7-1d84b3e70cb5 · inbound
Speculative Prefill: Turbocharging TTFT with Lightweight and Training-Free Token Importance Estimation CacheGen: KV Cache Compression and Streaming for Fast Large Language Model Serving
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e560a3a0-8f46-49e1-8f74-80956214fa5a · inbound
Cache-Craft: Managing Chunk-Caches for Efficient Retrieval-Augmented Generation CacheGen: KV Cache Compression and Streaming for Fast Large Language Model Serving
Reference 53
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d04da29a-e91d-4121-9b85-9c9c9a1bacee · inbound
SwiftSpec: Ultra-Low Latency LLM Decoding by Scaling Asynchronous Speculative Decoding CacheGen: KV Cache Compression and Streaming for Fast Large Language Model Serving
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8d66700f-ad39-4d58-a4f0-4ec86695dbfc · inbound
ViFusion: In-Network Tensor Fusion for Scalable Video Feature Indexing CacheGen: KV Cache Compression and Streaming for Fast Large Language Model Serving
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation aa5dc71d-861b-48c0-9a3a-05f55e5e1dc4 · inbound
On Evaluating Performance of LLM Inference Serving Systems CacheGen: KV Cache Compression and Streaming for Fast Large Language Model Serving
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 347e9ca1-f020-4eed-8969-5d0979cfd26f · inbound
MultiFluxAI Enhancing Platform Engineering with Advanced Agent-Orchestrated Retrieval Systems CacheGen: KV Cache Compression and Streaming for Fast Large Language Model Serving
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f388e6c0-2d7e-45fb-a46c-9d99ab23b9cf · inbound
Ubiquitous Intelligence Via Wireless Network-Driven LLMs Evolution CacheGen: KV Cache Compression and Streaming for Fast Large Language Model Serving
Reference 51
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 09e55922-d738-414a-9c2b-16ce9c3cedde · inbound
Adaptive KV Cache Reuse for Fast Long-Context LLM Serving CacheGen: KV Cache Compression and Streaming for Fast Large Language Model Serving
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 2694c557-0caa-45ea-8992-a1bfed5e6542 · inbound
QCFuse: Query-Aware Cache Fusion via Compressed View for Efficient RAG Serving CacheGen: KV Cache Compression and Streaming for Fast Large Language Model Serving
Reference 45
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation e221043e-7543-4001-a33a-60b3ebba5179 · inbound
Demystifying the Design Space and Best Practices for Heterogeneous LLM Inference and Serving CacheGen: KV Cache Compression and Streaming for Fast Large Language Model Serving
Reference 40
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation b66cc25a-7a9f-4cbf-b764-20fd0aedde85 · inbound
Demystifying the Design Space and Best Practices for Heterogeneous LLM Inference and Serving CacheGen: KV Cache Compression and Streaming for Fast Large Language Model Serving
Reference 40
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 0a8d3eba-7310-4e24-8daf-309770ebc037 · inbound
What to Keep, What to Forget: A Rate--Distortion View of Memory Compaction in LLMs and Agents CacheGen: KV Cache Compression and Streaming for Fast Large Language Model Serving
Reference 74
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation dcc0f714-5a6a-48f1-8a67-3630cf0085b9 · inbound
Smarter and Cheaper at Once: Byte-Exact KV-Cache Grafting Turns a Frozen Small Model into a Verified-Knowledge Flywheel CacheGen: KV Cache Compression and Streaming for Fast Large Language Model Serving
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2de3a19a-a0f9-4dc1-9201-8d2d0e35ce6d · inbound
Persistent Computational State: A Session-Centric Runtime for Generative World Models CacheGen: KV Cache Compression and Streaming for Fast Large Language Model Serving
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d8f82b89-0ba8-445c-93ef-2f2041318699 · inbound
Spatial Prefix Caching for Wireless Edge LLM Inference: A Stochastic-Geometry and Queueing Framework CacheGen: KV Cache Compression and Streaming for Fast Large Language Model Serving
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.