Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 9 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 29 inbound Pith citation observations for arXiv:2406.17565.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-09T04:32:01.202215Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-07-04T09:39:46.803577Z
0 of 0 outbound references displayed
External citation measurements
No source-named external measurement is stored.
No outbound reference observations are available for this paper version.
Observation 3b653b39-854e-4c09-bc2c-fb4c64f5f2fc · inbound
BatchLLM: Optimizing Large Batched LLM Inference with Global Prefix Sharing and Throughput-oriented Token Batching MemServe: Context Caching for Disaggregated LLM Serving with Elastic Memory Pool
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 08a01091-abe2-435f-a039-308069137136 · inbound
HACK: Homomorphic Acceleration via Compression of the Key-Value Cache for Disaggregated LLM Inference MemServe: Context Caching for Disaggregated LLM Serving with Elastic Memory Pool
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0b30afa7-a920-4a3a-92ee-b8aab4eef3cd · inbound
From Human Memory to AI Memory: A Survey on Memory Mechanisms in the Era of LLMs MemServe: Context Caching for Disaggregated LLM Serving with Elastic Memory Pool
Reference 122
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 943e6b8a-7059-48d4-b290-02b90832b38b · inbound
CoDec: Prefix-Shared Decoding Kernel for LLMs MemServe: Context Caching for Disaggregated LLM Serving with Elastic Memory Pool
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3f1b93ad-e1b0-450a-ac38-19ff0a11b44e · inbound
Beyond the Buzz: A Pragmatic Take on Inference Disaggregation MemServe: Context Caching for Disaggregated LLM Serving with Elastic Memory Pool
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 964a02ee-e3b0-467c-b77d-b3cd16ae2297 · inbound
DAM: Dynamic Attention Mask for Long-Context Large Language Model Inference Acceleration MemServe: Context Caching for Disaggregated LLM Serving with Elastic Memory Pool
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cb7ffe8d-b83e-4d3b-9d8f-7859cb9a33bf · inbound
Rethinking Caching for LLM Serving Systems: Beyond Traditional Heuristics MemServe: Context Caching for Disaggregated LLM Serving with Elastic Memory Pool
Reference 2024
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3b62d94d-a129-4d69-9b26-5a56356e42ca · inbound
Taming the Chaos: Coordinated Autoscaling for Heterogeneous and Disaggregated LLM Inference MemServe: Context Caching for Disaggregated LLM Serving with Elastic Memory Pool
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 26143fc2-cb77-4548-87f0-3d769c800d34 · inbound
Secure Multi-LLM Agentic AI and Agentification for Edge General Intelligence by Zero-Trust: A Survey MemServe: Context Caching for Disaggregated LLM Serving with Elastic Memory Pool
Reference 121
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 26a19a2d-3907-4e66-92b0-bba1b5cab032 · inbound
Boosting Embodied AI Agents through Perception-Generation Disaggregation and Asynchronous Pipeline Execution MemServe: Context Caching for Disaggregated LLM Serving with Elastic Memory Pool
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ae877024-34f8-43a3-9549-b818ea3d9c92 · inbound
DuetServe: Harmonizing Prefill and Decode for LLM Serving via Adaptive GPU Multiplexing MemServe: Context Caching for Disaggregated LLM Serving with Elastic Memory Pool
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4066d4c7-8537-4b24-91ea-a2cc0bb9a66a · inbound
SuperInfer: SLO-Aware Rotary Scheduling and Memory Management for LLM Inference on Superchips MemServe: Context Caching for Disaggregated LLM Serving with Elastic Memory Pool
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 47bd091b-30b8-416a-a99f-bbff14be2f8f · inbound
Efficient Remote KV Cache Reuse with GPU-native Video Codec MemServe: Context Caching for Disaggregated LLM Serving with Elastic Memory Pool
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation da2b1e96-3bd8-47b1-81d1-29da03e930ce · inbound
Towards Efficient Large Vision-Language Models: A Comprehensive Survey on Inference Strategies MemServe: Context Caching for Disaggregated LLM Serving with Elastic Memory Pool
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 3101e9b5-02e8-46ba-8f2e-d1bca161f339 · inbound
InfiniLoRA: Disaggregated Multi-LoRA Serving for Large Language Models MemServe: Context Caching for Disaggregated LLM Serving with Elastic Memory Pool
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation fd481473-3656-4800-8567-ddac1ac91067 · inbound
Taming Request Imbalance: SLO-Aware Scheduling for Disaggregated LLM Inference MemServe: Context Caching for Disaggregated LLM Serving with Elastic Memory Pool
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation aef8e12f-c4e3-4c25-9137-6f58e5ec7722 · inbound
Taming Request Imbalance: SLO-Aware Scheduling for Disaggregated LLM Inference MemServe: Context Caching for Disaggregated LLM Serving with Elastic Memory Pool
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation e463e356-0d2e-4706-94b3-2b2a0c63b35a · inbound
ObjectCache: Layerwise Object-Storage Retrieval for KV Cache Reuse MemServe: Context Caching for Disaggregated LLM Serving with Elastic Memory Pool
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation e820ca28-003b-4c8e-a73c-6b4fa10c1ba3 · inbound
AlignedServe: Orchestrating Prefix-aware Batching to Build a High-throughput and Computing-efficient LLM Serving System MemServe: Context Caching for Disaggregated LLM Serving with Elastic Memory Pool
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 9c02aad4-b481-4ca1-877e-d54c2f24c980 · inbound
Idleness is Relative: Exploiting Tool-Call Idle Windows for Offloading in Agentic Systems with MORI MemServe: Context Caching for Disaggregated LLM Serving with Elastic Memory Pool
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 2fd66820-e86d-4091-940f-115992c883d0 · inbound
Leyline: KV Cache Directives for Agentic Inference MemServe: Context Caching for Disaggregated LLM Serving with Elastic Memory Pool
Reference 42
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 116ba9ea-3c2a-4e3b-8b0d-22aa9632bda0 · inbound
ASAP: A Disaggregated and Asynchronous Inference System for MoE Prefill MemServe: Context Caching for Disaggregated LLM Serving with Elastic Memory Pool
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation ef9aefc8-c8cb-4722-b9f5-3ab05fb142b0 · inbound
KernelFlume: Elastic Core-Attention Scaling for Agentic Long-Context Decoding MemServe: Context Caching for Disaggregated LLM Serving with Elastic Memory Pool
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation fd7edb7a-853e-4df5-93bf-7fe941c9f6ca · inbound
From Tensor Buffer to Distributed Memory Hierarchy: A Survey of KV Cache Management for LLM Serving MemServe: Context Caching for Disaggregated LLM Serving with Elastic Memory Pool
Reference 47
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9bd133b5-8c0d-498d-af8e-8c50ce925c61 · inbound
[AAFLOW+] Stateful Operator Abstraction with Zero-Copy Distributed KV Cache Orchestration for Multi-Agent Workflows MemServe: Context Caching for Disaggregated LLM Serving with Elastic Memory Pool
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 98f1939f-4852-44c6-ad80-d2c249913367 · inbound
DualDecoder: Accelerate Long Context LLM Inference by Predictive Prefetch MemServe: Context Caching for Disaggregated LLM Serving with Elastic Memory Pool
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c63fd332-4148-4190-a355-98bd631e8bc3 · inbound
Rethinking AI Cloud Infrastructure for Agentic Serving Systems with the Aries Experimentation Framework MemServe: Context Caching for Disaggregated LLM Serving with Elastic Memory Pool
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 251d395c-dede-4ccd-bd4c-523d8ccf8a65 · inbound
Spatial Prefix Caching for Wireless Edge LLM Inference: A Stochastic-Geometry and Queueing Framework MemServe: Context Caching for Disaggregated LLM Serving with Elastic Memory Pool
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2b17282e-2cbb-4913-896d-4eaa09d08b5c · inbound
Energy-Efficient LLM Serving via Disaggregated Attention--FFN and Flexible Frequency Scaling MemServe: Context Caching for Disaggregated LLM Serving with Elastic Memory Pool
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.