Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 8 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 6 inbound Pith citation observations for arXiv:2412.04964.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-05T17:55:07.138663Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-05-16T09:57:43.035594Z
0 of 0 outbound references displayed
External citation measurements
No source-named external measurement is stored.
No outbound reference observations are available for this paper version.
Observation cdaf4a06-f029-4ae2-8f28-43a37df27a65 · inbound
TPLA: Tensor Parallel Latent Attention for Efficient Disaggregated Prefill and Decode Inference Flash Communication: Reducing Tensor Parallelization Bottleneck for Fast Large Language Model Inference
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8a2df785-e639-439d-b2dd-6cafd657326f · inbound
HAP: Hybrid Adaptive Parallelism for Efficient Mixture-of-Experts Inference Flash Communication: Reducing Tensor Parallelization Bottleneck for Fast Large Language Model Inference
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7f515dee-8e2a-4c60-8b9a-b2a0fd5eae2c · inbound
Analytical Provisioning for Attention-FFN Disaggregated LLM Serving under Stochastic Workloads Flash Communication: Reducing Tensor Parallelization Bottleneck for Fast Large Language Model Inference
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation beac9573-1959-47b8-a92f-361c852472b4 · inbound
A Switch-Centric In-Network Architecture for Accelerating LLM Inference in Shared-Memory Network Flash Communication: Reducing Tensor Parallelization Bottleneck for Fast Large Language Model Inference
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 7e236abc-011e-492c-8e5a-1b6b6084bfcb · inbound
CommFuse: Hiding Tail Latency via Communication Decomposition and Fusion for Distributed LLM Training Flash Communication: Reducing Tensor Parallelization Bottleneck for Fast Large Language Model Inference
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation ab26d428-582b-45e6-9500-1169c2860d4f · inbound
TACO: Efficient Communication Compression of Intermediate Tensors for Scalable Tensor-Parallel LLM Training Flash Communication: Reducing Tensor Parallelization Bottleneck for Fast Large Language Model Inference
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.