Pith. sign in

Paper Citation Record · LEDGER

LayerKV: Optimizing Large Language Model Serving with Layer-wise KV Cache Management

As of 5 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 6 inbound Pith citation observations for arXiv:2410.00428.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2410.00428 v3

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 6 of 6 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-05T06:32:48.257954+00:00

measured 6 of 6 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-01T21:09:10.825037Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-06-30T17:24:56.959321Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation bca2aa03-58fd-45a2-adcc-31e71317e029 · inbound

Efficient Remote KV Cache Reuse with GPU-native Video Codec cites this paper.

Efficient Remote KV Cache Reuse with GPU-native Video Codec LayerKV: Optimizing Large Language Model Serving with Layer-wise KV Cache Management

Reference 70

Resolution
verified exact
arxiv_id, observed 2026-05-16T05:22:22.824323Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-16T05:21:04.555356Z digest=sha256:5a70a2a159546f4883715e9d7c4885a0c07c9037ff6c304ff41a06759be8a187

Observation 0b7c41ff-21a0-42bd-b047-83131df9db7a · inbound

ForkKV: Scaling Multi-LoRA Agent Serving via Copy-on-Write Disaggregated KV Cache cites this paper.

ForkKV: Scaling Multi-LoRA Agent Serving via Copy-on-Write Disaggregated KV Cache LayerKV: Optimizing Large Language Model Serving with Layer-wise KV Cache Management

Reference 65

Resolution
verified exact
arxiv_id, observed 2026-05-11T05:10:54.903875Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T18:16:49.292491Z digest=sha256:5c2b93575a7aa95827acf2b80f80e9294f86cba1780b68c12d4da2b73c091216

Observation 92e32207-aca9-4950-81da-403be932bc1d · inbound

Adaptive KV Cache Reuse for Fast Long-Context LLM Serving cites this paper.

Adaptive KV Cache Reuse for Fast Long-Context LLM Serving LayerKV: Optimizing Large Language Model Serving with Layer-wise KV Cache Management

Reference 36

Resolution
verified exact
arxiv_id, observed 2026-06-30T17:24:56.960758Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-06-30T17:23:13.154458Z digest=sha256:6200c07f8144befd9ba859276730018a1b23f4af4f8b08b8d3fb6fe0389109c5

Observation baea7759-e46b-49ec-9f53-6230a91772e4 · inbound

PagedWeight: Efficient MoE LLM Serving with Dynamic Quality-Aware Weight Quantization cites this paper.

PagedWeight: Efficient MoE LLM Serving with Dynamic Quality-Aware Weight Quantization LayerKV: Optimizing Large Language Model Serving with Layer-wise KV Cache Management

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-01T21:09:10.825037Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T21:09:10.825037Z digest=sha256:d0de93727bf0b2fdb3729649ade51c1087268a5474bfe61b7c8b8f8b18c5d521

Observation 2a1b2494-8811-4498-b14e-4b0f6b4fe1db · inbound

Beyond Storage: State as a Runtime Control Problem in Parallel and Distributed Systems cites this paper.

Beyond Storage: State as a Runtime Control Problem in Parallel and Distributed Systems LayerKV: Optimizing Large Language Model Serving with Layer-wise KV Cache Management

Reference 230

Resolution
unresolved
no resolver link, observed 2026-08-01T19:51:22.631756Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T19:51:22.631756Z digest=sha256:4d64797ee3ca04b6e11babd17647ae46add48a80f616867e57236a108c5dc89f

Observation d91cb980-1b46-4f98-b2e4-c768f8990f3f · inbound

Persistent Computational State: A Session-Centric Runtime for Generative World Models cites this paper.

Persistent Computational State: A Session-Centric Runtime for Generative World Models LayerKV: Optimizing Large Language Model Serving with Layer-wise KV Cache Management

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-01T07:46:27.035461Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T07:46:27.035461Z digest=sha256:a2e9a3a05b6a994abf984a376604289d2cd17007157b1e52ca2d0bd80f715a02