Pith. sign in

Paper Citation Record · LEDGER

EPIC: Efficient Position-Independent Caching for Serving Large Language Models

As of 22 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 16 inbound Pith citation observations for arXiv:2410.15332.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2410.15332 v3

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 16 of 16 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-21T06:32:19.484+00:00

measured 16 of 16 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-16T05:49:09.832470Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

2
arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 16eb82f5-604a-43fe-aeef-c27b535607fa · inbound

Auditing Prompt Caching in Language Model APIs cites this paper.

Auditing Prompt Caching in Language Model APIs EPIC: Efficient Position-Independent Caching for Serving Large Language Models

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-08T11:39:40.694233Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T11:39:40.694233Z digest=sha256:7aeaf7d9e2c3d2bb91c2679bd67038964906af8066d5eed62fddd97b4564239b

Observation 8931695c-fd7c-4b42-b8a1-570a3e36c545 · inbound

From Human Memory to AI Memory: A Survey on Memory Mechanisms in the Era of LLMs cites this paper.

From Human Memory to AI Memory: A Survey on Memory Mechanisms in the Era of LLMs EPIC: Efficient Position-Independent Caching for Serving Large Language Models

Reference 136

Resolution
verified exact
arxiv_id, observed 2026-05-17T11:05:09.882503Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-17T11:05:09.588491Z digest=sha256:ea0aaf921535ca5944815e4a628b556d27c848827b1368646b8dad3815cd2296

Observation cdf5b360-90f8-4283-a5a1-e226183b5cf6 · inbound

Taming the Titans: A Survey of Efficient LLM Inference Serving cites this paper.

Taming the Titans: A Survey of Efficient LLM Inference Serving EPIC: Efficient Position-Independent Caching for Serving Large Language Models

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-16T05:49:09.832470Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T05:49:09.832470Z digest=sha256:5ae0af333c3179a30297087f59c2991a9283c5f245adc3563e7d0af21254f265

Observation 928ea50b-3914-431e-ae90-2c0002094ad7 · inbound

Semantic Caching of Contextual Summaries for Efficient Question-Answering with Language Models cites this paper.

Semantic Caching of Contextual Summaries for Efficient Question-Answering with Language Models EPIC: Efficient Position-Independent Caching for Serving Large Language Models

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-15T20:58:47.910040Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:58:47.910040Z digest=sha256:55a80e817b0aa4333d8142182a74bce938b3e003a4d0c320119582c2d3131775

Observation a5740179-dd6a-415a-98d5-908d4a52a845 · inbound

InfoFlow KV: Information-Flow-Aware KV Recomputation for Long Context cites this paper.

InfoFlow KV: Information-Flow-Aware KV Recomputation for Long Context EPIC: Efficient Position-Independent Caching for Serving Large Language Models

Reference 2020

Resolution
unresolved
no resolver link, observed 2026-08-02T18:47:14.923342Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T18:47:14.923342Z digest=sha256:ec5edd57dd35a35a6d27e951c2806055ef66a57c02b03d81c4c3e9483700f427

Observation 34912571-7245-46c1-b0b2-3bd55e22af32 · inbound

Parallelizing Tool Execution and LLM Generation for Low-Latency Agent Serving cites this paper.

Parallelizing Tool Execution and LLM Generation for Low-Latency Agent Serving EPIC: Efficient Position-Independent Caching for Serving Large Language Models

Reference 18

Resolution
unresolved
no resolver link, observed 2026-07-13T22:20:51.838309Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-13T22:20:51.838309Z digest=sha256:508ace545dd244d2d062e11be65762f539eda2a873e2765d7fe11d0049373409

Observation 268afca9-41b5-4291-abf5-297d99f65b59 · inbound

Low-Scaling Many-Body Green's Function Calculations for Molecular Systems via Interacting-Bath Dynamical Embedding Theory cites this paper.

Low-Scaling Many-Body Green's Function Calculations for Molecular Systems via Interacting-Bath Dynamical Embedding Theory EPIC: Efficient Position-Independent Caching for Serving Large Language Models

Reference 13

Resolution
unresolved
no resolver link, observed 2026-07-13T13:32:41.086180Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-13T13:32:41.086180Z digest=sha256:435a16fbbd7fee4cff0962295cdc2de7451d229050e9c65873052bc8c2d0a786

Observation 6fdbffdf-911b-4a0f-8e44-ab042ef400de · inbound

TokenDance: Scaling Multi-Agent LLM Serving via Collective KV Cache Sharing cites this paper.

TokenDance: Scaling Multi-Agent LLM Serving via Collective KV Cache Sharing EPIC: Efficient Position-Independent Caching for Serving Large Language Models

Reference 13

Resolution
verified exact
arxiv_id, observed 2026-05-13T18:03:04.298579Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-13T18:02:46.266111Z digest=sha256:0c701c99e1c3ac153497d80813caabdda1da881db144e4e207ef1b0d68a2a4f8

Observation 4acad4f8-620e-4e68-9a0e-771e8987acdd · inbound

CachePrune: Privacy-Aware and Fine-Grained KV Cache Sharing for Efficient LLM Inference cites this paper.

CachePrune: Privacy-Aware and Fine-Grained KV Cache Sharing for Efficient LLM Inference EPIC: Efficient Position-Independent Caching for Serving Large Language Models

Reference 17

Resolution
verified exact
arxiv_id, observed 2026-05-25T04:15:19.801774Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-25T04:14:44.387842Z digest=sha256:4a540eb07a11c8216693acc7c414258cfb4b9cc10c689b307d9cdb6467c0da55

Observation a04edc82-e6e5-4ba9-9e96-a1d6c05e310f · inbound

Grounded Cache Routing for Retrieval-Augmented Generation: When Is It Safe to Reuse an Answer? cites this paper.

Grounded Cache Routing for Retrieval-Augmented Generation: When Is It Safe to Reuse an Answer? EPIC: Efficient Position-Independent Caching for Serving Large Language Models

Reference 6

Resolution
verified exact
arxiv_id, observed 2026-06-29T17:23:45.167477Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-06-29T17:16:48.588593Z digest=sha256:20b6911169aecb9e9d5abbf501f715dc8750bb0701413636c5356d6a70695265

Observation cdd55039-99b2-4a1d-92b0-6f07db62d2a0 · inbound

QCFuse: Query-Aware Cache Fusion via Compressed View for Efficient RAG Serving cites this paper.

QCFuse: Query-Aware Cache Fusion via Compressed View for Efficient RAG Serving EPIC: Efficient Position-Independent Caching for Serving Large Language Models

Reference 24

Resolution
verified exact
arxiv_id, observed 2026-07-02T12:46:57.448784Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-06-28T01:47:43.240850Z digest=sha256:cc10e6e69a28a9ff79c7af152ce5981b65722a2d5780facc7e02631b3bf07ca8

Observation 4db2cd42-8f24-4528-a461-c7255f05848d · inbound

SIFT: Selective-Index For Fast Compute of RAG Prefill by Exploiting Attention Invariance cites this paper.

SIFT: Selective-Index For Fast Compute of RAG Prefill by Exploiting Attention Invariance EPIC: Efficient Position-Independent Caching for Serving Large Language Models

Reference 12

Resolution
verified exact
arxiv_id, observed 2026-07-03T01:37:31.265369Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-06-27T16:24:31.109508Z digest=sha256:d24166954ab10dba2c01c742d04d8e2b8af0d54db5cdb20070ed13b3ec87726d

Observation c3beeea0-ffa3-42c7-886c-5708c787b198 · inbound

Functional Cache Grafting: Robust and Rapid Code-Policy Synthesis for Embodied Agents cites this paper.

Functional Cache Grafting: Robust and Rapid Code-Policy Synthesis for Embodied Agents EPIC: Efficient Position-Independent Caching for Serving Large Language Models

Reference 10

Resolution
verified exact
arxiv_id, observed 2026-06-27T05:30:35.811260Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-06-27T05:26:48.431739Z digest=sha256:c6c5c061c1a2b0d761ed3c159a5db11dbbaefa04773457c9405cc33843d6311e

Observation 0fc35fc9-d6c3-4e0b-a9d0-2722d0cda3ce · inbound

Functional Cache Grafting: Robust and Rapid Code-Policy Synthesis for Embodied Agents cites this paper.

Functional Cache Grafting: Robust and Rapid Code-Policy Synthesis for Embodied Agents EPIC: Efficient Position-Independent Caching for Serving Large Language Models

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-02T11:44:03.503592Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T11:44:03.503592Z digest=sha256:13e3b00e47b185dcd518fd30679fe72d118cc7233d3bfe661e29d69fcc424de8

Observation 7087d409-6f77-41e8-bc6d-2c1eb698c4bd · inbound

MiniPIC: Flexible Position-Independent Caching in <100LOC cites this paper.

MiniPIC: Flexible Position-Independent Caching in <100LOC EPIC: Efficient Position-Independent Caching for Serving Large Language Models

Reference 3

Resolution
verified exact
arxiv_id, observed 2026-06-27T07:20:42.213077Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-06-27T07:12:45.449658Z digest=sha256:3fc4067b428833b7abb6928d50b09ee2a0f3916512ae8953850beffbf361b394

Observation 334edc70-9177-4c44-a5f5-ab7899638126 · inbound

Compute Globally, Materialize Locally: The Memory Contract of Sparse Event-KV cites this paper.

Compute Globally, Materialize Locally: The Memory Contract of Sparse Event-KV EPIC: Efficient Position-Independent Caching for Serving Large Language Models

Reference 6

Resolution
unresolved
no resolver link, observed 2026-07-30T15:53:12.075501Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-30T15:53:12.075501Z digest=sha256:befd9c199a04633fe7ec03feae31e88233f67e9e4c05cb0e910f2af6d4b00dc4