Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 7 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 23 inbound Pith citation observations for arXiv:2401.11181.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-07T10:26:32.736179Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z
0 of 0 outbound references displayed
External citation measurements
6
arxiv_reference, observed 2026-08-05T02:28:24.338817Z
No outbound reference observations are available for this paper version.
Observation cd8da1e9-bc32-448d-bd9d-4a7ffbce1000 · inbound
A Survey on Efficient Inference for Large Language Models Inference without Interference: Disaggregate LLM Inference for Mixed Downstream Workloads
Reference 273
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation d539e8c3-45f1-43ce-a2b0-9593364440c7 · inbound
Beyond the Buzz: A Pragmatic Take on Inference Disaggregation Inference without Interference: Disaggregate LLM Inference for Mixed Downstream Workloads
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5f148349-ce14-4b64-8106-9106552df42e · inbound
Nexus:Proactive Intra-GPU Disaggregation of Prefill and Decode in LLM Serving Inference without Interference: Disaggregate LLM Inference for Mixed Downstream Workloads
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f33f8277-8b6a-4991-b6f4-b43908aa318b · inbound
Taming the Chaos: Coordinated Autoscaling for Heterogeneous and Disaggregated LLM Inference Inference without Interference: Disaggregate LLM Inference for Mixed Downstream Workloads
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1f96ef1a-0a70-4f5b-96ed-3cd4aff33490 · inbound
Hetis: Serving LLMs in Heterogeneous GPU Clusters with Fine-grained and Dynamic Parallelism Inference without Interference: Disaggregate LLM Inference for Mixed Downstream Workloads
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 26ad8cfa-5076-4b7b-b8bd-67eb2edd7fbc · inbound
FlexPipe: Adapting Dynamic LLM Serving Through Inflight Pipeline Refactoring in Fragmented Serverless Clusters Inference without Interference: Disaggregate LLM Inference for Mixed Downstream Workloads
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 1a795543-9eae-427b-8f11-15ae5f4ad8c2 · inbound
STAR: Decode-Phase Rescheduling for LLM Inference Inference without Interference: Disaggregate LLM Inference for Mixed Downstream Workloads
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 0af25593-fd88-4ed5-954b-b2bef949192e · inbound
ODMA: On-Demand Memory Allocation Strategy for LLM Serving on LPDDR-Class Accelerators Inference without Interference: Disaggregate LLM Inference for Mixed Downstream Workloads
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation d20fdeaf-85f2-4752-af64-7f07217c58e6 · inbound
Efficient Multi-round LLM Inference over Disaggregated Serving Inference without Interference: Disaggregate LLM Inference for Mixed Downstream Workloads
Reference 1994
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1ce18c33-284a-4fde-b632-ee06b6f0368e · inbound
Coral: Cost-Efficient Multi-LLM Serving over Heterogeneous Cloud GPUs Inference without Interference: Disaggregate LLM Inference for Mixed Downstream Workloads
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 193c6fde-95d1-4e3c-bdc9-5cfc75a66375 · inbound
VeriCache: Turning Lossy KV Cache into Lossless LLM Inference Inference without Interference: Disaggregate LLM Inference for Mixed Downstream Workloads
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 93fbeaed-f880-4c22-88d8-a296d194c9da · inbound
AlignedServe: Orchestrating Prefix-aware Batching to Build a High-throughput and Computing-efficient LLM Serving System Inference without Interference: Disaggregate LLM Inference for Mixed Downstream Workloads
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 34fb4efc-43e1-4098-ae80-ff6ca91927d7 · inbound
Human-Less LLM Serving: Quantifying the Human Tax on Throughput Inference without Interference: Disaggregate LLM Inference for Mixed Downstream Workloads
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 66c5f74e-e132-435a-8a50-d755ca4e86e2 · inbound
Geometry-Aware Online Scheduling for LLM Serving: From Theoretical Bound to System Practice Inference without Interference: Disaggregate LLM Inference for Mixed Downstream Workloads
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 91619bb5-87eb-472e-8b08-6cf731256fcb · inbound
Demystifying the Design Space and Best Practices for Heterogeneous LLM Inference and Serving Inference without Interference: Disaggregate LLM Inference for Mixed Downstream Workloads
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 04af9445-3e67-4eb8-b5ea-8d59b1fe8549 · inbound
Demystifying the Design Space and Best Practices for Heterogeneous LLM Inference and Serving Inference without Interference: Disaggregate LLM Inference for Mixed Downstream Workloads
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 9045a234-d428-4ed8-947a-61dad41f26e5 · inbound
Towards Load-Aware Prefill Deflection for Disaggregated LLM Serving Inference without Interference: Disaggregate LLM Inference for Mixed Downstream Workloads
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation ac815cd7-9513-4cb1-b790-b1cb655a2309 · inbound
CoCoScale: Leveraging Layer-wise Scaling to Unlock the Potential of Online LLM Serving Inference without Interference: Disaggregate LLM Inference for Mixed Downstream Workloads
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ebab7f57-b3e2-4028-9376-3b53766a92b9 · inbound
Sangam: Efficiently Serving Diffusion LLMs with the AR Stack Inference without Interference: Disaggregate LLM Inference for Mixed Downstream Workloads
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 25516814-d561-4c3a-aa0a-af61bd4072d0 · inbound
AutoSLO: Practical Latency SLOs on Cloud Data Warehouses -- Extended Version Inference without Interference: Disaggregate LLM Inference for Mixed Downstream Workloads
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 88f9298f-b660-4f3d-8b42-61c58cd035e9 · inbound
SmartGen: Seamless Disaggregated LLM Inference with Selective KV Cache Transfer Inference without Interference: Disaggregate LLM Inference for Mixed Downstream Workloads
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b49bd378-5512-4d7b-b8f6-08a197292084 · inbound
Rethinking AI Cloud Infrastructure for Agentic Serving Systems with the Aries Experimentation Framework Inference without Interference: Disaggregate LLM Inference for Mixed Downstream Workloads
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7d2fc877-c3e9-4a5b-9681-8eb28e071b5f · inbound
Energy-Efficient LLM Serving via Disaggregated Attention--FFN and Flexible Frequency Scaling Inference without Interference: Disaggregate LLM Inference for Mixed Downstream Workloads
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.