Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 8 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 10 inbound Pith citation observations for arXiv:2405.10637.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-06T13:57:15.793887Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-07-01T23:16:23.532286Z
0 of 0 outbound references displayed
External citation measurements
No source-named external measurement is stored.
No outbound reference observations are available for this paper version.
Observation 8facd824-f426-4216-a2f3-95d1e0c5a236 · inbound
When Attention Sink Emerges in Language Models: An Empirical View Layer-Condensed KV Cache for Efficient Inference of Large Language Models
Reference 53
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation c79d223c-ceef-4c26-8a6d-3c70c232ac82 · inbound
Semantic Integrity Matters: Benchmarking and Preserving High-Density Reasoning in KV Cache Compression Layer-Condensed KV Cache for Efficient Inference of Large Language Models
Reference 57
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 7b6f0485-9572-4977-8526-ce7db81918e2 · inbound
CaliDrop: KV Cache Compression with Calibration Layer-Condensed KV Cache for Efficient Inference of Large Language Models
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 60be5ed8-16ce-4a7e-941f-1cd3df3493bf · inbound
TPLA: Tensor Parallel Latent Attention for Efficient Disaggregated Prefill and Decode Inference Layer-Condensed KV Cache for Efficient Inference of Large Language Models
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ec10e0a1-1afa-4399-bb3f-7967d322121b · inbound
CacheTrap: Unveiling a Stealthier Gray-Box Trojan against LLMs Layer-Condensed KV Cache for Efficient Inference of Large Language Models
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 78a08397-e062-44bb-aad8-398ec5914b0f · inbound
ClusterFusion++: Expanding Cluster-Level Fusion to Full Transformer-Block Decoding Layer-Condensed KV Cache for Efficient Inference of Large Language Models
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 66c5a0d8-90cb-4d5a-b40c-6265b75558e4 · inbound
Do Value Vectors in Deep Layers Need Context from the Residual Stream? Layer-Condensed KV Cache for Efficient Inference of Large Language Models
Reference 106
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation a5ab00fb-e0d4-41e5-a9d3-70dc5eccf3e1 · inbound
Do Value Vectors in Deep Layers Need Context from the Residual Stream? Layer-Condensed KV Cache for Efficient Inference of Large Language Models
Reference 46
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ccc29096-3f45-49e3-85f1-84b97ec275aa · inbound
Structured Thoughts For Improved Reasoning And Context Pruning Layer-Condensed KV Cache for Efficient Inference of Large Language Models
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation dc2c4eee-5961-41b9-800a-1342369c9d8b · inbound
Understanding Is Done Early: A Depth Division of Labor in Large Language Models and Its Use for Unbounded-Context Memory Layer-Condensed KV Cache for Efficient Inference of Large Language Models
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.