Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-04T23:39:09.503875Z
Paper Citation Record · LEDGER
As of 17 August 2026, this Paper Citation Record lists 21 of 21 outbound references and 0 inbound Pith citation observations for arXiv:2608.01657.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-04T23:39:09.503875Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-17T06:30:58.91139+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
21 of 21 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 387b4860-b496-471e-819e-ee815ad404f2 · outbound
Preserving Admission Responsibility in Multi-Tenant Large Language Model Prefix Caches Efficient memory management for large language model serving with PagedAttention,
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 8970443a-57a6-494f-b853-f566639fa1f5 · outbound
Preserving Admission Responsibility in Multi-Tenant Large Language Model Prefix Caches SGLang: Efficient execution of structured language model programs,
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 475d677f-11ba-4d0f-bdf3-d532130e5e82 · outbound
Preserving Admission Responsibility in Multi-Tenant Large Language Model Prefix Caches Stateful large language model serving with Pensieve,
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation ec41eb9c-d281-446c-a7e8-5441c48b0ba4 · outbound
Preserving Admission Responsibility in Multi-Tenant Large Language Model Prefix Caches Online context caching for distributed large language models serving,
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation b438ddf6-5854-4404-873f-7603880d0a9b · outbound
Preserving Admission Responsibility in Multi-Tenant Large Language Model Prefix Caches MELL: Memory-efficient large language model serving via multi-GPU KV cache management,
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation ac31c024-14e6-427f-8c54-89d9bd68a6cc · outbound
Preserving Admission Responsibility in Multi-Tenant Large Language Model Prefix Caches Mooncake: Trading more storage for less computation—a KVCache-centric architecture for serving LLM chatbot,
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation a628dc5c-0bd3-4956-85f9-a1a1ac216f76 · outbound
Preserving Admission Responsibility in Multi-Tenant Large Language Model Prefix Caches KVCache cache in the wild: Characterizing and optimizing KVCache cache at a large cloud provider,
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 5c1e0539-4df6-4216-9724-927c2e65c6cc · outbound
Preserving Admission Responsibility in Multi-Tenant Large Language Model Prefix Caches Fairness in serving large language models,
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 31e2e63d-7eda-42e8-91fe-a42e51674042 · outbound
Preserving Admission Responsibility in Multi-Tenant Large Language Model Prefix Caches FIFO queues are all you need for cache eviction,
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 2393ba6f-3e14-4d71-8e14-73a252c42dc4 · outbound
Preserving Admission Responsibility in Multi-Tenant Large Language Model Prefix Caches FairRide: Near- optimal, fair cache sharing,
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 34140d50-9b9f-47c3-bac6-4390299b0e00 · outbound
Preserving Admission Responsibility in Multi-Tenant Large Language Model Prefix Caches NyxCache: Flexible and efficient multi-tenant persistent memory caching,
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 7cd55043-5d03-49bd-934d-f70664315653 · outbound
Preserving Admission Responsibility in Multi-Tenant Large Language Model Prefix Caches TinyLFU: A highly efficient cache admission policy,
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation e52f807e-2837-49e9-a41a-94b75d625a4d · outbound
Preserving Admission Responsibility in Multi-Tenant Large Language Model Prefix Caches CacheBlend: Fast large language model serving for RAG with cached knowledge fusion,
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 9843a93a-fc54-491c-a553-03b251c23e44 · outbound
Preserving Admission Responsibility in Multi-Tenant Large Language Model Prefix Caches Oneiros: KV cache optimization through parameter remapping for multi-tenant LLM serving,
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation e80bb4ab-9b9d-4c1f-9d3e-bc2eaa8f2c3e · outbound
Preserving Admission Responsibility in Multi-Tenant Large Language Model Prefix Caches ServeGen: Workload characterization and generation of large language model serving in production,
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation f7b9beaf-771a-439a-be53-4e5b1b3dec39 · outbound
Preserving Admission Responsibility in Multi-Tenant Large Language Model Prefix Caches IC-Cache: Efficient large language model serving via in-context caching,
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 75cb134f-501c-4873-a8b6-340314b8e29b · outbound
Preserving Admission Responsibility in Multi-Tenant Large Language Model Prefix Caches CacheGen: KV cache compression and streaming for fast large language model serving,
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 7bab2b8f-4ad4-4afd-8566-958bb6a29a01 · outbound
Preserving Admission Responsibility in Multi-Tenant Large Language Model Prefix Caches ThunderServe: High-performance and cost-efficient LLM serving in cloud environments,
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation ce5eac51-163c-412e-94ac-9437cde73904 · outbound
Preserving Admission Responsibility in Multi-Tenant Large Language Model Prefix Caches Jenga: Effective memory management for serving LLM with heterogeneity,
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 370e30f8-fc6a-476c-a98b-35f1b71f4d59 · outbound
Preserving Admission Responsibility in Multi-Tenant Large Language Model Prefix Caches LServe: Efficient long-sequence LLM serving with unified sparse attention,
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 7a66cbfc-27fd-4516-9ab4-d4402cda02e8 · outbound
Preserving Admission Responsibility in Multi-Tenant Large Language Model Prefix Caches Memory- efficient KV cache optimization for large language model inference at the edge,
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
No inbound Pith citation observations are available.