Pith. sign in

Paper Citation Record · LEDGER

LUMEN: Coordinated Failure Recovery for Distributed LLM Serving

As of 18 August 2026, this Paper Citation Record lists 45 of 45 outbound references and 0 inbound Pith citation observations for arXiv:2606.17787.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2606.17787 v1

Coverage vector

measured 45 of 45 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-06-26T22:53:58.352008Z

measured 45 of 45 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-18T06:34:40.430872+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

45 of 45 outbound references displayed

  • verified exact5
  • verified fuzzy0
  • unresolved40
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation a50ef0c8-068d-4fa3-9305-9fbaff6a797e · outbound

This paper cites https:// docs.vllm.ai/en/stable/deployment/k8s/, 2026.

LUMEN: Coordinated Failure Recovery for Distributed LLM Serving https:// docs.vllm.ai/en/stable/deployment/k8s/, 2026

Reference 1

Resolution
unresolved
no resolver link, observed 2026-06-26T22:53:58.352008Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-26T22:53:58.352008Z digest=sha256:78f9bb2ac5e5ff8625e1916c6dade7a866bc8ac3606ca1a981d3ed7764e9ef21

Observation df5d1cf0-a1c7-4c6d-b9b1-43563d9ff6c8 · outbound

This paper cites https: //github.com/flashinfer-ai/flashinfer, 2026.

LUMEN: Coordinated Failure Recovery for Distributed LLM Serving https: //github.com/flashinfer-ai/flashinfer, 2026

Reference 2

Resolution
unresolved
no resolver link, observed 2026-06-26T22:53:58.352008Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-26T22:53:58.352008Z digest=sha256:5ff2de15611e6058a7c3219ea3c58118f04ebfd893742e10ffa2be6e1f8d9e23

Observation 3eec8535-1c49-4233-8c49-e7445cde2f5c · outbound

This paper cites an unresolved cited work.

LUMEN: Coordinated Failure Recovery for Distributed LLM Serving Unresolved cited work

Reference 3

Resolution
unresolved
no resolver link, observed 2026-06-26T22:53:58.352008Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-26T22:53:58.352008Z digest=sha256:7eb17003856318706da015df1b56ca1d0808b695f94f938b61b0ac0af766e45f

Observation acf3725c-eac4-47f5-94b1-c5d9e7be79c4 · outbound

This paper cites https: //huggingface.co/docs/text-generation- inference, 2026.

LUMEN: Coordinated Failure Recovery for Distributed LLM Serving https: //huggingface.co/docs/text-generation- inference, 2026

Reference 4

Resolution
unresolved
no resolver link, observed 2026-06-26T22:53:58.352008Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-26T22:53:58.352008Z digest=sha256:43a440e6b7f6be4878138a7352ca1d750a27976b3a20d06fbba9ca2b8c31c180

Observation 701ca8b4-f9b7-4b55-8585-2de165840038 · outbound

This paper cites https://github.com/kserve/kserve, 2026.

LUMEN: Coordinated Failure Recovery for Distributed LLM Serving https://github.com/kserve/kserve, 2026

Reference 5

Resolution
unresolved
no resolver link, observed 2026-06-26T22:53:58.352008Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-26T22:53:58.352008Z digest=sha256:f86469b1a1418b330f42e9db41d405a6d42d7445c1a67a216f245fdacc92df58

Observation 23f45e03-73f7-413e-a239-e7085c7510c8 · outbound

This paper cites https: //github.com/triton-inference-server/server, 2026.

LUMEN: Coordinated Failure Recovery for Distributed LLM Serving https: //github.com/triton-inference-server/server, 2026

Reference 6

Resolution
unresolved
no resolver link, observed 2026-06-26T22:53:58.352008Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-26T22:53:58.352008Z digest=sha256:0e4c0b2b6c428d846f4ec5a715f6853f75936eb5b00927e7fa0016cacf691abd

Observation c4695858-47cc-434c-a8e0-c99a4a748883 · outbound

This paper cites https://pytorch.org, 2026.

LUMEN: Coordinated Failure Recovery for Distributed LLM Serving https://pytorch.org, 2026

Reference 7

Resolution
unresolved
no resolver link, observed 2026-06-26T22:53:58.352008Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-26T22:53:58.352008Z digest=sha256:6963f5b33ae76e68500fb0025a0479b128887887e529dab677aba749ec67015c

Observation afd8ff5c-6475-49f8-9b2f-9e36f1cd873e · outbound

This paper cites https://zeromq.org, 2026.

LUMEN: Coordinated Failure Recovery for Distributed LLM Serving https://zeromq.org, 2026

Reference 8

Resolution
unresolved
no resolver link, observed 2026-06-26T22:53:58.352008Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-26T22:53:58.352008Z digest=sha256:bc41d343593f0a85a3c599ed4a28e5e58194dd0a1c9336e8ddcb87927a56f127

Observation b713c946-798f-4d39-aad2-cdf79a4cd3a1 · outbound

This paper cites Vidur: A large-scale simulation framework for LLM inference.

LUMEN: Coordinated Failure Recovery for Distributed LLM Serving Vidur: A large-scale simulation framework for LLM inference

Reference 9

Resolution
unresolved
no resolver link, observed 2026-06-26T22:53:58.352008Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-26T22:53:58.352008Z digest=sha256:76064082a0e53e0ae1d53eb99becf66f8a65163f3dcaa6fef2723cc8c2402df3

Observation cf4fbf55-0dfd-4729-8e6a-f573a0dda2e6 · outbound

This paper cites Taming throughput- latency tradeoff in LLM inference with Sarathi-Serve.

LUMEN: Coordinated Failure Recovery for Distributed LLM Serving Taming throughput- latency tradeoff in LLM inference with Sarathi-Serve

Reference 10

Resolution
unresolved
no resolver link, observed 2026-06-26T22:53:58.352008Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-26T22:53:58.352008Z digest=sha256:4fd536e4a52c2373e0942baaa182ad3412848a6bc8ead76c9ae452b275da8f89

Observation 31f6ed29-fa9a-4e33-98c0-5ede73efd04a · outbound

This paper cites an unresolved cited work.

LUMEN: Coordinated Failure Recovery for Distributed LLM Serving Unresolved cited work

Reference 11

Resolution
unresolved
no resolver link, observed 2026-06-26T22:53:58.352008Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-26T22:53:58.352008Z digest=sha256:21138c0a2c5f9afa428451caec3ce786b30a91a0df3b161983c9f62a74a60179

Observation 98d1a7b1-350b-44ed-9c07-be3785c4ef8b · outbound

This paper cites Accelerating Large Language Model Decoding with Speculative Sampling.

LUMEN: Coordinated Failure Recovery for Distributed LLM Serving Accelerating Large Language Model Decoding with Speculative Sampling

Reference 12

Resolution
verified exact
local_arxiv, observed 2026-07-03T23:09:01.189168Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-06-26T22:53:58.352008Z digest=sha256:7596445ec62cea78d56f32b9a2fb3964e427f50193cd393ef101a423b310a21a

Observation 8d98aca0-3450-4bb5-8e64-ca52f26e3ecd · outbound

This paper cites Recycle: Resilient training of large DNNs using pipeline adaptation.

LUMEN: Coordinated Failure Recovery for Distributed LLM Serving Recycle: Resilient training of large DNNs using pipeline adaptation

Reference 13

Resolution
unresolved
no resolver link, observed 2026-06-26T22:53:58.352008Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-26T22:53:58.352008Z digest=sha256:a04f9ecde04232c660f9e9459139669e228855760a7a911df1a5215c4c324287

Observation 94ea7503-a455-4be8-9f72-d01228cb1099 · outbound

This paper cites Cost-efficient large language model serving for multi-turn conversations with CachedAtten- tion.

LUMEN: Coordinated Failure Recovery for Distributed LLM Serving Cost-efficient large language model serving for multi-turn conversations with CachedAtten- tion

Reference 14

Resolution
unresolved
no resolver link, observed 2026-06-26T22:53:58.352008Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-26T22:53:58.352008Z digest=sha256:701dc57c807f2edd3e9a46c8610bd7aa20fe46c65cf2fe57282988e4ee23c71d

Observation a747b482-a093-44e9-b235-fc9c34e81dfb · outbound

This paper cites Characterization of large language model devel- opment in the datacenter.

LUMEN: Coordinated Failure Recovery for Distributed LLM Serving Characterization of large language model devel- opment in the datacenter

Reference 15

Resolution
unresolved
no resolver link, observed 2026-06-26T22:53:58.352008Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-26T22:53:58.352008Z digest=sha256:abc285b3566ee43e629159fb7c12f0024e7e3af322b3d0ad7482bc55b034b515

Observation 56d1b946-fa90-4ad8-ad8f-bfb6e3f6d005 · outbound

This paper cites Oobleck: Resilient distributed training of large models using pipeline templates.

LUMEN: Coordinated Failure Recovery for Distributed LLM Serving Oobleck: Resilient distributed training of large models using pipeline templates

Reference 16

Resolution
unresolved
no resolver link, observed 2026-06-26T22:53:58.352008Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-26T22:53:58.352008Z digest=sha256:4ea1c0277d35b0232ef64eeffeefdb5a16367e58b0424ca8a7afa9e22d98ef76

Observation 515ecadb-84bc-437c-8c44-7a9e3072bdeb · outbound

This paper cites GhostServe: A lightweight checkpointing system in the shadow for fault-tolerant LLM serving.

LUMEN: Coordinated Failure Recovery for Distributed LLM Serving GhostServe: A lightweight checkpointing system in the shadow for fault-tolerant LLM serving

Reference 17

Resolution
unresolved
no resolver link, observed 2026-06-26T22:53:58.352008Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-26T22:53:58.352008Z digest=sha256:ed2d235728c2ad00c59fdaccd005323ada63b5107db238acd8c06032e4501d34

Observation 95689ca0-8f85-440a-8936-2ef05fadd90c · outbound

This paper cites MegaScale: Scaling large language model training to more than 10,000 GPUs.

LUMEN: Coordinated Failure Recovery for Distributed LLM Serving MegaScale: Scaling large language model training to more than 10,000 GPUs

Reference 18

Resolution
unresolved
no resolver link, observed 2026-06-26T22:53:58.352008Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-26T22:53:58.352008Z digest=sha256:e9ada7206c54561b767199c3cd024cc477cb56515c4b08aef979fc76059ed008

Observation ed395e94-2832-470a-ad0d-f0f5027ebac4 · outbound

This paper cites Revisiting reliability in large-scale machine learning research clusters.

LUMEN: Coordinated Failure Recovery for Distributed LLM Serving Revisiting reliability in large-scale machine learning research clusters

Reference 19

Resolution
unresolved
no resolver link, observed 2026-06-26T22:53:58.352008Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-26T22:53:58.352008Z digest=sha256:2dee7b63ec978f89fae6914cc394b71ba897667ab4fb54cb7f1a91e87008fd46

Observation 8bc8901c-819e-41fb-ba2a-8d082e8e96a7 · outbound

This paper cites Efficient memory man- agement for large language model serving with Page- dAttention.

LUMEN: Coordinated Failure Recovery for Distributed LLM Serving Efficient memory man- agement for large language model serving with Page- dAttention

Reference 20

Resolution
unresolved
no resolver link, observed 2026-06-26T22:53:58.352008Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-26T22:53:58.352008Z digest=sha256:d2fa2dd3bfca21a9d1dc95c443621f2f94700860b9cdc459344a385cec0cbdfa

Observation 5175a2b5-3a68-4f8b-8a1d-cfa9fdadb63e · outbound

This paper cites Fast inference from transformers via speculative decoding.

LUMEN: Coordinated Failure Recovery for Distributed LLM Serving Fast inference from transformers via speculative decoding

Reference 21

Resolution
unresolved
no resolver link, observed 2026-06-26T22:53:58.352008Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-26T22:53:58.352008Z digest=sha256:1d10361ee0753c35d89317eebf0cfe56502cc2f06c1a5b399297f14c02d17630

Observation added2ae-d8c7-463b-9ed1-25ef75b9edc5 · outbound

This paper cites PEARL: Parallel speculative decoding with adaptive draft length.

LUMEN: Coordinated Failure Recovery for Distributed LLM Serving PEARL: Parallel speculative decoding with adaptive draft length

Reference 22

Resolution
unresolved
no resolver link, observed 2026-06-26T22:53:58.352008Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-26T22:53:58.352008Z digest=sha256:a43b0d8433065e25bbad18035a2262008507f8d9137837e30ab4825f66abb0f1

Observation eb39cfe3-cd5f-4adb-87eb-ef22e76cc1ee · outbound

This paper cites CacheGen: KV cache compression and streaming for fast large language model serving.

LUMEN: Coordinated Failure Recovery for Distributed LLM Serving CacheGen: KV cache compression and streaming for fast large language model serving

Reference 23

Resolution
unresolved
no resolver link, observed 2026-06-26T22:53:58.352008Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-26T22:53:58.352008Z digest=sha256:0ee608a004022f0ac40e414c1a7d3e0c2063b2205c1712b75b61167603bfecce

Observation 5bac4c3d-c534-47ac-855a-ef8d24467246 · outbound

This paper cites The Llama 3 Herd of Models.

LUMEN: Coordinated Failure Recovery for Distributed LLM Serving The Llama 3 Herd of Models

Reference 24

Resolution
verified exact
local_arxiv, observed 2026-07-03T23:09:01.186199Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-06-26T22:53:58.352008Z digest=sha256:3dde1df7c3401c5c7c38b3111ece751b2a2952bb6baa2a33ea3ee5a06101cedb

Observation dba359e2-87a3-45fd-a6a3-4205127ea8ab · outbound

This paper cites AMUSD: Asynchronous multi- device speculative decoding for LLM acceleration.

LUMEN: Coordinated Failure Recovery for Distributed LLM Serving AMUSD: Asynchronous multi- device speculative decoding for LLM acceleration

Reference 25

Resolution
unresolved
no resolver link, observed 2026-06-26T22:53:58.352008Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-26T22:53:58.352008Z digest=sha256:ced2e10425939f36f57ef26178c7aabe3c38230bd433201d70224d570ad85ae5

Observation f31f6144-a051-47f3-a1be-f18954711ad6 · outbound

This paper cites Splitwise: Efficient generative LLM inference using phase splitting.

LUMEN: Coordinated Failure Recovery for Distributed LLM Serving Splitwise: Efficient generative LLM inference using phase splitting

Reference 26

Resolution
unresolved
no resolver link, observed 2026-06-26T22:53:58.352008Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-26T22:53:58.352008Z digest=sha256:f31baf922e617657076727cb560c93d997228e636692f8e0c1816feb3acfce65

Observation 19009165-96f7-46d6-9df1-04dc08ef38c7 · outbound

This paper cites ECCheck: Enhancing in-memory check- point with erasure coding in distributed DNN training.

LUMEN: Coordinated Failure Recovery for Distributed LLM Serving ECCheck: Enhancing in-memory check- point with erasure coding in distributed DNN training

Reference 27

Resolution
unresolved
no resolver link, observed 2026-06-26T22:53:58.352008Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-26T22:53:58.352008Z digest=sha256:78c45adf067f47f0f0eec8bf31b2ce824e060aebafa05cddf191580b8e95d100

Observation 4ad37d1b-49c6-43b4-a423-a497bfc98a1b · outbound

This paper cites Towards resiliency in large language model serving with KevlarFlow.

LUMEN: Coordinated Failure Recovery for Distributed LLM Serving Towards resiliency in large language model serving with KevlarFlow

Reference 28

Resolution
verified exact
arxiv_id, observed 2026-07-03T23:09:01.193291Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-06-26T22:53:58.352008Z digest=sha256:1ca3083932789f1ef218faf6f75810d2ce0413a9acc6ef95a055b45c4d467240

Observation 153437a6-363a-4fe3-88a0-09f5421cd3e1 · outbound

This paper cites Mooncake: Trading more storage for less computation – a KVCache-centric architecture for serving LLM chat- bot.

LUMEN: Coordinated Failure Recovery for Distributed LLM Serving Mooncake: Trading more storage for less computation – a KVCache-centric architecture for serving LLM chat- bot

Reference 29

Resolution
unresolved
no resolver link, observed 2026-06-26T22:53:58.352008Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-26T22:53:58.352008Z digest=sha256:4bacefd045733f7b3d2c7942b6c262cb79e813962639f77825df7a9e3688f06b

Observation 360ecb2f-7a53-46a8-aebe-021cd4b73805 · outbound

This paper cites ShareGPT conversation dataset.

LUMEN: Coordinated Failure Recovery for Distributed LLM Serving ShareGPT conversation dataset

Reference 30

Resolution
unresolved
no resolver link, observed 2026-06-26T22:53:58.352008Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-26T22:53:58.352008Z digest=sha256:e0884c154d3a288a74a9b652a7313a26944560c751604d8765ad75ac94cfbd71

Observation 5fc473c8-2533-4ddb-9c19-f87ba7c755f4 · outbound

This paper cites FlexGen: High- throughput generative inference of large language mod- els with a single GPU.

LUMEN: Coordinated Failure Recovery for Distributed LLM Serving FlexGen: High- throughput generative inference of large language mod- els with a single GPU

Reference 31

Resolution
unresolved
no resolver link, observed 2026-06-26T22:53:58.352008Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-26T22:53:58.352008Z digest=sha256:bb72516ed2db38319104400d831f2ac7c4a4d8497f20e704fb0477b3f84396a5

Observation 0befe7a2-ac5b-4a55-a9cc-07357fb3f436 · outbound

This paper cites DéjàVu: KV-cache streaming for fast, fault-tolerant generative LLM serv- ing.

LUMEN: Coordinated Failure Recovery for Distributed LLM Serving DéjàVu: KV-cache streaming for fast, fault-tolerant generative LLM serv- ing

Reference 32

Resolution
unresolved
no resolver link, observed 2026-06-26T22:53:58.352008Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-26T22:53:58.352008Z digest=sha256:7253bfef05762464fdc7a1b890692bb7cd94726d45c966b5c2b8473b628a731a

Observation 94d765e1-aebf-4d92-b496-5c9e849f8c99 · outbound

This paper cites Llumnix: Dy- namic scheduling for large language model serving.

LUMEN: Coordinated Failure Recovery for Distributed LLM Serving Llumnix: Dy- namic scheduling for large language model serving

Reference 33

Resolution
unresolved
no resolver link, observed 2026-06-26T22:53:58.352008Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-26T22:53:58.352008Z digest=sha256:8fd1a1d068950b3d84780f74be32d2353515e682d99abba8f729db1ac5884f36

Observation fb7d7f18-bf5b-48ad-b556-301b453006ff · outbound

This paper cites Bamboo: Making preemptible in- stances resilient for affordable training of large DNNs.

LUMEN: Coordinated Failure Recovery for Distributed LLM Serving Bamboo: Making preemptible in- stances resilient for affordable training of large DNNs

Reference 34

Resolution
unresolved
no resolver link, observed 2026-06-26T22:53:58.352008Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-26T22:53:58.352008Z digest=sha256:2ec4caab75c529de432f557bb79a6b5dd6b8cd4a3eb56fad6d59bb86b16a7c50

Observation e352b913-5f7a-4493-8a1f-644d3e34d3d3 · outbound

This paper cites Llama 2: Open Foundation and Fine-Tuned Chat Models.

LUMEN: Coordinated Failure Recovery for Distributed LLM Serving Llama 2: Open Foundation and Fine-Tuned Chat Models

Reference 35

Resolution
verified exact
local_arxiv, observed 2026-07-03T23:09:01.195699Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-06-26T22:53:58.352008Z digest=sha256:b515b0633a47b15700d85adbdd2d8d138f9b6528e53a57a21a03a229dd43477d

Observation 60346195-0523-4335-bf39-8b4459dc9add · outbound

This paper cites Gomez, Lukasz Kaiser, and Illia Polosukhin.

LUMEN: Coordinated Failure Recovery for Distributed LLM Serving Gomez, Lukasz Kaiser, and Illia Polosukhin

Reference 36

Resolution
unresolved
no resolver link, observed 2026-06-26T22:53:58.352008Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-26T22:53:58.352008Z digest=sha256:23b25cfed3a754ed6b7836aa84d967b057a77fd1bb117d1d4e04d24385c822fa

Observation 3b86e601-9fd8-484d-879f-9d759b9a2487 · outbound

This paper cites ByteCheckpoint: A unified checkpointing system for large foundation model development.

LUMEN: Coordinated Failure Recovery for Distributed LLM Serving ByteCheckpoint: A unified checkpointing system for large foundation model development

Reference 37

Resolution
unresolved
no resolver link, observed 2026-06-26T22:53:58.352008Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-26T22:53:58.352008Z digest=sha256:07edf16210267ddf882dddedaef3729b8853a245ee8b086cf9d5ab077b100a4c

Observation 767c681d-bcb1-4ac1-aa04-ad2c3c248818 · outbound

This paper cites an unresolved cited work.

LUMEN: Coordinated Failure Recovery for Distributed LLM Serving Unresolved cited work

Reference 38

Resolution
unresolved
no resolver link, observed 2026-06-26T22:53:58.352008Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-26T22:53:58.352008Z digest=sha256:5d49886f11eace5e15bcac5a3ffc8eae34a69eeac3a06629564e39e98392f748

Observation d92a333d-db58-44b8-a9e6-6bed58a0a42a · outbound

This paper cites Fast distributed inference serving for large language models.

LUMEN: Coordinated Failure Recovery for Distributed LLM Serving Fast distributed inference serving for large language models

Reference 39

Resolution
unresolved
no resolver link, observed 2026-06-26T22:53:58.352008Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-26T22:53:58.352008Z digest=sha256:96daadc55d85eac269a6a1aea01059560b81c9ac01ffd9706e19f0c0d2f15e86

Observation 6d92c734-89a3-4b37-8d4f-b355a1466328 · outbound

This paper cites FailSafe: High-performance resilient serv- ing.

LUMEN: Coordinated Failure Recovery for Distributed LLM Serving FailSafe: High-performance resilient serv- ing

Reference 40

Resolution
unresolved
no resolver link, observed 2026-06-26T22:53:58.352008Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-26T22:53:58.352008Z digest=sha256:e48881aa4197a6224b61942b5bfb6ffdee8dd7ae1356e79e999fc8789117a36d

Observation 132556ef-0f08-4109-b294-bf595c950621 · outbound

This paper cites Qwen3 Technical Report.

LUMEN: Coordinated Failure Recovery for Distributed LLM Serving Qwen3 Technical Report

Reference 41

Resolution
verified exact
local_arxiv, observed 2026-07-03T23:09:01.191928Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-06-26T22:53:58.352008Z digest=sha256:40a03c21a841a4f875e7c79351f48c257af8cb45f1e898db3bd59664b0fa2ae1

Observation 147a22fe-75dd-4747-8ed3-2c5113223e8e · outbound

This paper cites Orca: A distributed serving system for transformer-based generative models.

LUMEN: Coordinated Failure Recovery for Distributed LLM Serving Orca: A distributed serving system for transformer-based generative models

Reference 42

Resolution
unresolved
no resolver link, observed 2026-06-26T22:53:58.352008Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-26T22:53:58.352008Z digest=sha256:fb2dc858ccd09714d0de3d9d095d656dbd9295b4cb373325f66b327285172676

Observation 7c57243e-6044-475c-a09e-84826308776b · outbound

This paper cites Gonzalez, Clark Bar- rett, and Ying Sheng.

LUMEN: Coordinated Failure Recovery for Distributed LLM Serving Gonzalez, Clark Bar- rett, and Ying Sheng

Reference 43

Resolution
unresolved
no resolver link, observed 2026-06-26T22:53:58.352008Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-26T22:53:58.352008Z digest=sha256:2c66b136878a95570e1712d0ce6e5acaa56306e6b6f9536ac0b956e438f42bdf

Observation a2349f40-b36b-4963-bab5-575d3f837a11 · outbound

This paper cites Dist- Serve: Disaggregating prefill and decoding for goodput- optimized large language model serving.

LUMEN: Coordinated Failure Recovery for Distributed LLM Serving Dist- Serve: Disaggregating prefill and decoding for goodput- optimized large language model serving

Reference 44

Resolution
unresolved
no resolver link, observed 2026-06-26T22:53:58.352008Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-26T22:53:58.352008Z digest=sha256:12d2b309371f6e01ffc54c51720a2250a27a629c642804d35dfdfa70fef56cb9

Observation 650d2676-44dd-4bc8-8308-af8833cd27e8 · outbound

This paper cites Resiliency at scale: Managing Google’s TPUv4 machine learning supercomputer.

LUMEN: Coordinated Failure Recovery for Distributed LLM Serving Resiliency at scale: Managing Google’s TPUv4 machine learning supercomputer

Reference 45

Resolution
unresolved
no resolver link, observed 2026-06-26T22:53:58.352008Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-26T22:53:58.352008Z digest=sha256:382fab2cf65772b12f37376ea8a2811c0e40e8c4f34628d4a352c0c022553299

Pith citing papers

No inbound Pith citation observations are available.