Pith. sign in

Paper Citation Record · LEDGER

GhostServe: A Lightweight Checkpointing System in the Shadow for Fault-Tolerant LLM Serving

As of 5 August 2026, this Paper Citation Record lists 24 of 24 outbound references and 0 inbound Pith citation observations for arXiv:2605.00831.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2605.00831 v1

Coverage vector

measured 24 of 24 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-05-15T00:37:08.671539Z

measured 24 of 24 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-05T06:32:48.257954+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

24 of 24 outbound references displayed

  • verified exact15
  • verified fuzzy5
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch4

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 1f32c318-668d-4544-9bcd-c918306c2685 · outbound

This paper cites SARATHI: Efficient LLM Inference by Piggybacking Decodes with Chunked Prefills.

GhostServe: A Lightweight Checkpointing System in the Shadow for Fault-Tolerant LLM Serving SARATHI: Efficient LLM Inference by Piggybacking Decodes with Chunked Prefills

Reference 1

Resolution
verified exact
arxiv_id, observed 2026-05-16T06:31:47.535686Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-15T00:37:08.671539Z digest=sha256:307bf1ae10211dfe6e5de28dbee1191f91dd1073755bb551c36899cb809707dc

Observation 40cf2f55-f826-430b-b478-0bb215411241 · outbound

This paper cites 2025.No Request Left Behind: Tackling Heterogeneity in Long-Context LLM Inference with Medha.

GhostServe: A Lightweight Checkpointing System in the Shadow for Fault-Tolerant LLM Serving 2025.No Request Left Behind: Tackling Heterogeneity in Long-Context LLM Inference with Medha

Reference 2

Resolution
verified exact
arxiv_id, observed 2026-05-15T00:38:23.090087Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-15T00:37:08.671539Z digest=sha256:07ac8c5337d43a437d2a44d4c696570de27ab56a4aca98e1fb21b59163f31e85

Observation 4f336952-cfe7-480a-9cf2-fc50bb54a60a · outbound

This paper cites K., Janakiraman, R., and Xu, L.

GhostServe: A Lightweight Checkpointing System in the Shadow for Fault-Tolerant LLM Serving K., Janakiraman, R., and Xu, L

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-05-15T00:38:23.691453Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-15T00:37:08.671539Z digest=sha256:96782b5652012b73af11087aebeca9ba24e47b6c6333a43d27591ee0654597a6

Observation f189bdce-0ca6-4c60-a276-1464c4d55027 · outbound

This paper cites D., Dhariwal, P., Neelakantan, A., Shyam, P., Sastry, G., Askell, A., et al.

GhostServe: A Lightweight Checkpointing System in the Shadow for Fault-Tolerant LLM Serving D., Dhariwal, P., Neelakantan, A., Shyam, P., Sastry, G., Askell, A., et al

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-05-15T00:38:23.693289Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-15T00:37:08.671539Z digest=sha256:bd39eb50c66247dde48bddcedcb7256d556d9cc7f38fe51b5fe19101081dcfd1

Observation 03632557-c26b-4c89-a771-071cdba79daf · outbound

This paper cites LithOS: An Operating System for Efficient Machine Learning on GPUs.

GhostServe: A Lightweight Checkpointing System in the Shadow for Fault-Tolerant LLM Serving LithOS: An Operating System for Efficient Machine Learning on GPUs

Reference 5

Resolution
verified exact
arxiv_id, observed 2026-05-15T00:38:23.080048Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-15T00:37:08.671539Z digest=sha256:b237bca831ff2c8d216580942e1d76cf76cb06f68e1097130959fc3b7261ac03

Observation ba6c971b-44b2-4d65-9383-188482e59c92 · outbound

This paper cites FlashAttention-2: Faster Attention with Better Parallelism and Work Partitioning.

GhostServe: A Lightweight Checkpointing System in the Shadow for Fault-Tolerant LLM Serving FlashAttention-2: Faster Attention with Better Parallelism and Work Partitioning

Reference 6

Resolution
verified exact
local_arxiv, observed 2026-05-15T00:38:23.086652Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-15T00:37:08.671539Z digest=sha256:20a886925babc0f3b58b260eb06f098b567eecdc9d4aeab1a548665943eb5494

Observation 7fbb0742-452e-4d74-a18e-27b87f258025 · outbound

This paper cites G., and Yang, J.

GhostServe: A Lightweight Checkpointing System in the Shadow for Fault-Tolerant LLM Serving G., and Yang, J

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-05-15T00:38:23.695085Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-15T00:37:08.671539Z digest=sha256:c4a7081327cf3f2b8428524bbc0d315c85e03151aa3ef13e254925e7697f7445

Observation 5e329564-c35a-48a8-a297-97cca508cf6b · outbound

This paper cites DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning.

GhostServe: A Lightweight Checkpointing System in the Shadow for Fault-Tolerant LLM Serving DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning

Reference 8

Resolution
metadata mismatch
local_arxiv, observed 2026-05-15T00:38:23.137446Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-15T00:37:08.671539Z digest=sha256:ba617167e1612d72e2a09fdb46e3b2b084b810d7f0adda3da03bb3a1e1870e27

Observation 5f1a7cbd-251c-4db1-9a9d-330552fe9476 · outbound

This paper cites Capacity-Aware Inference: Mitigating the Straggler Effect in Mixture of Experts.

GhostServe: A Lightweight Checkpointing System in the Shadow for Fault-Tolerant LLM Serving Capacity-Aware Inference: Mitigating the Straggler Effect in Mixture of Experts

Reference 9

Resolution
verified exact
local_arxiv, observed 2026-05-15T00:38:23.083416Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-15T00:37:08.671539Z digest=sha256:de902fff7f25bafd8f9712849d919ee5300038a4cbfe460cfef62c6879503d99

Observation 1115cdd8-ca82-4100-bfc7-8462b890676d · outbound

This paper cites Training Compute-Optimal Large Language Models.

GhostServe: A Lightweight Checkpointing System in the Shadow for Fault-Tolerant LLM Serving Training Compute-Optimal Large Language Models

Reference 10

Resolution
metadata mismatch
local_arxiv, observed 2026-05-15T00:38:23.114487Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-15T00:37:08.671539Z digest=sha256:6b379c4d803a313b20885edc1048ccce2a562b4ac7002fec9d4afa48a61b6bca

Observation 50cd5564-0d7d-473b-9401-79f63bd87680 · outbound

This paper cites Demystifying Cost-Efficiency in LLM Serving over Heterogeneous GPUs.

GhostServe: A Lightweight Checkpointing System in the Shadow for Fault-Tolerant LLM Serving Demystifying Cost-Efficiency in LLM Serving over Heterogeneous GPUs

Reference 11

Resolution
verified exact
arxiv_id, observed 2026-05-15T00:38:23.108123Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-15T00:37:08.671539Z digest=sha256:7d1bc46d047318632be2360f40607e8b287756036da086d000044f195e0b3040

Observation bb503c89-5bcb-4c9e-91b2-9976bb34ac6f · outbound

This paper cites Scaling Laws for Neural Language Models.

GhostServe: A Lightweight Checkpointing System in the Shadow for Fault-Tolerant LLM Serving Scaling Laws for Neural Language Models

Reference 12

Resolution
verified exact
local_arxiv, observed 2026-05-15T00:38:23.100128Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-15T00:37:08.671539Z digest=sha256:849295c3ebf9cffb0b0fdf4509e4110f7053f10ad64aa8b8fc9fe8dd35fe40d9

Observation 7221dabb-4205-4175-9750-75eb80846d88 · outbound

This paper cites Revisiting reliability in large-scale machine learning research clusters.2025 IEEE International Symposium on High Performance Computer Architecture (HPCA), pp.

GhostServe: A Lightweight Checkpointing System in the Shadow for Fault-Tolerant LLM Serving Revisiting reliability in large-scale machine learning research clusters.2025 IEEE International Symposium on High Performance Computer Architecture (HPCA), pp

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-05-15T00:38:23.687864Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-15T00:37:08.671539Z digest=sha256:1531dfce41a5434be2300535d81ec533c28cd8b70dba95764944c932dcce80a2

Observation d3ff529d-b3d9-46e5-8ea7-b8a05c71a8b9 · outbound

This paper cites Understanding Stragglers in Large Model Training Using What-if Analysis.

GhostServe: A Lightweight Checkpointing System in the Shadow for Fault-Tolerant LLM Serving Understanding Stragglers in Large Model Training Using What-if Analysis

Reference 14

Resolution
verified exact
arxiv_id, observed 2026-05-15T00:38:23.118021Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-15T00:37:08.671539Z digest=sha256:fa33a82b824531e12d5ad702021f5b70f7023b72613cb801670d24cfeacd65a2

Observation 889ee8a7-05b6-4da5-8069-68d56240f472 · outbound

This paper cites Enhancing relia- bility in ai inference services: An empirical study on real production incidents.ArXiv, abs/2511.07424.

GhostServe: A Lightweight Checkpointing System in the Shadow for Fault-Tolerant LLM Serving Enhancing relia- bility in ai inference services: An empirical study on real production incidents.ArXiv, abs/2511.07424

Reference 15

Resolution
verified exact
arxiv_id, observed 2026-05-15T00:38:23.104043Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-15T00:37:08.671539Z digest=sha256:fbab373fefb68d9140667d0cc81cf8cc659add9047174ad356fb1b8da23c63b1

Observation 1283926b-249c-4c8a-a9d8-13a88c0ba556 · outbound

This paper cites Salpekar, R.

GhostServe: A Lightweight Checkpointing System in the Shadow for Fault-Tolerant LLM Serving Salpekar, R

Reference 16

Resolution
metadata mismatch
arxiv_id, observed 2026-05-15T00:38:23.077148Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-15T00:37:08.671539Z digest=sha256:649ffc2dd80a5cc99f5656c1d5510d175bb5bfc86b24b6d10c4df8ff16cdcc96

Observation 3d6f1dc7-4be7-4ede-a732-8a8f007b9148 · outbound

This paper cites Megatron-LM: Training Multi-Billion Parameter Language Models Using Model Parallelism.

GhostServe: A Lightweight Checkpointing System in the Shadow for Fault-Tolerant LLM Serving Megatron-LM: Training Multi-Billion Parameter Language Models Using Model Parallelism

Reference 17

Resolution
verified exact
local_arxiv, observed 2026-05-15T00:38:23.130518Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-15T00:37:08.671539Z digest=sha256:f8986f9c7436e0c22bc759d77b17a87b8c2b3aa8fcd75f4eb34c2543e973c0b3

Observation 854b60c6-4938-49f3-a83d-f527498fb772 · outbound

This paper cites 2024 crowdstrike-related it outages.

GhostServe: A Lightweight Checkpointing System in the Shadow for Fault-Tolerant LLM Serving 2024 crowdstrike-related it outages

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-05-15T00:38:23.689769Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-15T00:37:08.671539Z digest=sha256:a9b13d74246088124aa5e5ffe589a08b3741d210ff260a5598a97b4c287bf4fc

Observation e9d8b662-e73d-445e-86ed-15b90c7c7276 · outbound

This paper cites HuggingFace's Transformers: State-of-the-art Natural Language Processing.

GhostServe: A Lightweight Checkpointing System in the Shadow for Fault-Tolerant LLM Serving HuggingFace's Transformers: State-of-the-art Natural Language Processing

Reference 19

Resolution
verified exact
local_arxiv, observed 2026-05-15T00:38:23.127082Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-15T00:37:08.671539Z digest=sha256:47f12f3dc656e1933d61fd1f13b11fcd5693f9bc47917d7fefc761dc81aad019

Observation 1da15145-5a79-40c6-a3ca-a2837908f483 · outbound

This paper cites Fast Distributed Inference Serving for Large Language Models.

GhostServe: A Lightweight Checkpointing System in the Shadow for Fault-Tolerant LLM Serving Fast Distributed Inference Serving for Large Language Models

Reference 20

Resolution
verified exact
arxiv_id, observed 2026-05-17T11:23:46.844514Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-15T00:37:08.671539Z digest=sha256:3af657ff3338fe5e25fa1475fa04a64f96f9df6c484f1a126708687166acce37

Observation 5af68555-97f5-4986-ad77-b6ede239c494 · outbound

This paper cites Context Parallelism for Scalable Million-Token Inference.

GhostServe: A Lightweight Checkpointing System in the Shadow for Fault-Tolerant LLM Serving Context Parallelism for Scalable Million-Token Inference

Reference 21

Resolution
metadata mismatch
arxiv_id, observed 2026-05-15T00:38:23.124324Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-15T00:37:08.671539Z digest=sha256:c52eced68d464f9e02d5a450416b7ae829290a2952943e4cf6bd143a2fdc8740

Observation bffeda50-3fa6-4530-b5be-a9295d55c3e5 · outbound

This paper cites FlashInfer: Efficient and Customizable Attention Engine for LLM Inference Serving.

GhostServe: A Lightweight Checkpointing System in the Shadow for Fault-Tolerant LLM Serving FlashInfer: Efficient and Customizable Attention Engine for LLM Inference Serving

Reference 22

Resolution
verified exact
arxiv_id, observed 2026-05-16T13:26:34.840430Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-15T00:37:08.671539Z digest=sha256:6a333a047ac829086c58272633779740ecd2c82f1aceb10f5c9c18b9a8fd8512

Observation 6f8048e6-cdaa-4392-ac5a-1e1291b86f2a · outbound

This paper cites ALISA: Accelerating Large Language Model Inference via Sparsity-Aware KV Caching.

GhostServe: A Lightweight Checkpointing System in the Shadow for Fault-Tolerant LLM Serving ALISA: Accelerating Large Language Model Inference via Sparsity-Aware KV Caching

Reference 23

Resolution
verified exact
arxiv_id, observed 2026-05-15T00:38:23.111335Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-15T00:37:08.671539Z digest=sha256:a14584bb9119515eb4f25831a4b07258c9e33fc9b31a2b8540a6b58a2c6c170c

Observation fd98a068-3906-4903-9d7c-a58fd7ffb68b · outbound

This paper cites SampleAttention: Near-Lossless Acceleration of Long Context LLM Inference with Adaptive Structured Sparse Attention.

GhostServe: A Lightweight Checkpointing System in the Shadow for Fault-Tolerant LLM Serving SampleAttention: Near-Lossless Acceleration of Long Context LLM Inference with Adaptive Structured Sparse Attention

Reference 24

Resolution
verified exact
arxiv_id, observed 2026-05-15T00:38:23.097181Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-15T00:37:08.671539Z digest=sha256:0e34e2d9d2826fa8eaaa94b00f12c81324720b85c636fdff08dabed0a8ae0e01

Pith citing papers

No inbound Pith citation observations are available.