pith. sign in

Mooncake: A kvcache-centric disaggregated architecture for llm serving.ACM Transactions on Storage, 2024

2 Pith papers cite this work. Polarity classification is still indexing.

2 Pith papers citing it

fields

cs.DC 1 cs.LG 1

years

2026 2

representative citing papers

Prefill-as-a-Service: KVCache of Next-Generation Models Could Go Cross-Datacenter

cs.DC · 2026-04-16 · unverdicted · novelty 6.0

PrfaaS enables practical cross-datacenter prefill-decode disaggregation for hybrid-attention models via selective offloading, bandwidth-aware scheduling, and cache-aware placement, yielding 54% higher throughput and 64% lower P90 TTFT than homogeneous baselines in a 1T-parameter case study.

citing papers explorer

Showing 2 of 2 citing papers.