Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T13:22:45.314876Z
Paper Citation Record · LEDGER
As of 8 August 2026, this Paper Citation Record lists 13 of 13 outbound references and 0 inbound Pith citation observations for arXiv:2505.21919.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T13:22:45.314876Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
13 of 13 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 375ff02c-a8a4-408b-b728-6828018802f6 · outbound
Towards Efficient Key-Value Cache Management for Prefix Prefilling in LLM Inference CHIME: A Cache-Efficient and High-Performance Hybrid Index on Disaggregated Memory,
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 0799148e-4176-41c6-be15-11073ca27ef8 · outbound
Towards Efficient Key-Value Cache Management for Prefix Prefilling in LLM Inference Sherman: A Write-Optimized Distributed B+Tree Index on Disaggregated Memory,
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation f05966c6-ad4a-4800-b5df-a7844b415cc3 · outbound
Towards Efficient Key-Value Cache Management for Prefix Prefilling in LLM Inference Unlocking Longer Generation with Key-Value Cache Quantization
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation fc9c2d9e-0206-48e6-a911-c3301b27228b · outbound
Towards Efficient Key-Value Cache Management for Prefix Prefilling in LLM Inference vLLM vs TensorRT- LLM 12, Automatic Prefix Caching
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 9f28759a-2cf3-4dd8-8106-86b6acf5f8e9 · outbound
Towards Efficient Key-Value Cache Management for Prefix Prefilling in LLM Inference More Tokens, Lower Precision: Towards the Optimal Token-Precision Trade-off in KV Cache Compression
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ba612c96-afe4-4af9-9450-4a4d25d12738 · outbound
Towards Efficient Key-Value Cache Management for Prefix Prefilling in LLM Inference GQA: Training Generalized Multi-Query Transformer Models from Multi-Head Checkpoints
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6f72440f-ebc4-4f90-b435-07f7a5963d91 · outbound
Towards Efficient Key-Value Cache Management for Prefix Prefilling in LLM Inference Longformer: The long- document transformer,
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation d6e904c3-9c8b-42a7-95ff-0b9f54fb0e8d · outbound
Towards Efficient Key-Value Cache Management for Prefix Prefilling in LLM Inference Pie: Pooling CPU Memory for LLM Inference
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a97db814-edd7-46eb-84ce-9e6adf73d081 · outbound
Towards Efficient Key-Value Cache Management for Prefix Prefilling in LLM Inference Mooncake: Trading More Storage for Less Computation—A KVCache-centric Architecture for Serving LLM Chatbot,
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 0fb4072f-0a8b-49d0-9160-848e7435ae17 · outbound
Towards Efficient Key-Value Cache Management for Prefix Prefilling in LLM Inference CacheGen: KV Cache Compression and Streaming for Fast Large Language Model Serving,
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 5aeb4ef5-c030-4ff3-a977-6aa2a4c68dd7 · outbound
Towards Efficient Key-Value Cache Management for Prefix Prefilling in LLM Inference DeepSeek 3FS
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 75256a89-6565-4e20-b516-4c4f2b4578ab · outbound
Towards Efficient Key-Value Cache Management for Prefix Prefilling in LLM Inference IMPRESS: An Importance-Informed Multi-Tier Prefix KV Storage System for Large Language Model Inference,
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation e12c1413-4397-4f13-abee-b623222209de · outbound
Towards Efficient Key-Value Cache Management for Prefix Prefilling in LLM Inference Exploring cxl-based kv cache storage for llm serving,
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
No inbound Pith citation observations are available.