Pith. sign in

Paper Citation Record · LEDGER

BurstGPT: A Real-world Workload Dataset to Optimize LLM Serving Systems

As of 7 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 21 inbound Pith citation observations for arXiv:2401.17644.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2401.17644 v5

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 21 of 21 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 21 of 21 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-06T17:46:20.940919Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-07-07T20:34:09.759421Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 83f8b1b0-af64-4547-ae1d-9e31630141f0 · inbound

ServeGen: Workload Characterization and Generation of Large Language Model Serving in Production cites this paper.

ServeGen: Workload Characterization and Generation of Large Language Model Serving in Production BurstGPT: A Real-world Workload Dataset to Optimize LLM Serving Systems

Reference 45

Resolution
verified exact
arxiv_id, observed 2026-05-22T15:44:58.154778Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-22T15:42:05.266854Z digest=sha256:84d4f8ef5e83f4fe23c56bf1897c5e8d7a5706e1c4b322f2985122ed94061147

Observation b98b6759-a8c0-47e6-adb8-b97827578fab · inbound

Past-Future Scheduler for LLM Serving under SLA Guarantees cites this paper.

Past-Future Scheduler for LLM Serving under SLA Guarantees BurstGPT: A Real-world Workload Dataset to Optimize LLM Serving Systems

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-06T17:46:20.940919Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:46:20.940919Z digest=sha256:a36e55b2ee5616531b21e694529312f63db87bb7629fbfbb5ae3b315539d98e3

Observation c8f02fcc-0387-4b39-9431-998ed626a0a2 · inbound

Quality-of-Service Aware LLM Routing for Edge Computing with Multiple Experts cites this paper.

Quality-of-Service Aware LLM Routing for Edge Computing with Multiple Experts BurstGPT: A Real-world Workload Dataset to Optimize LLM Serving Systems

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-06T10:22:42.499001Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T10:22:42.499001Z digest=sha256:06cff72ef8eb09489e87b4e74e31d5166d99a1bdd2bf4c1a3544b4f9f827aa40

Observation c132c0cb-e3cd-44e3-8c06-33ebda30f0fa · inbound

Rethinking Caching for LLM Serving Systems: Beyond Traditional Heuristics cites this paper.

Rethinking Caching for LLM Serving Systems: Beyond Traditional Heuristics BurstGPT: A Real-world Workload Dataset to Optimize LLM Serving Systems

Reference 69

Resolution
unresolved
no resolver link, observed 2026-08-05T16:21:09.370531Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T16:21:09.370531Z digest=sha256:ec80d1648c821d14dc4f678a208f244a70c0aca3c3a8be742e19041b68b59d29

Observation cf06d08f-ec08-47af-8b66-4e8f0b46a461 · inbound

Hetis: Serving LLMs in Heterogeneous GPU Clusters with Fine-grained and Dynamic Parallelism cites this paper.

Hetis: Serving LLMs in Heterogeneous GPU Clusters with Fine-grained and Dynamic Parallelism BurstGPT: A Real-world Workload Dataset to Optimize LLM Serving Systems

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-04T20:53:33.420816Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:53:33.420816Z digest=sha256:33aab6ee36947e50a9f239ea0759872bebe3edec92a44437da8ff3f0a9fbf0b2

Observation de068e43-80e3-4e2f-a387-5f7c1b53a9c6 · inbound

WarmServe: Enabling One-for-Many GPU Prewarming for Multi-LLM Serving cites this paper.

WarmServe: Enabling One-for-Many GPU Prewarming for Multi-LLM Serving BurstGPT: A Real-world Workload Dataset to Optimize LLM Serving Systems

Reference 23

Resolution
verified exact
arxiv_id, observed 2026-05-22T12:24:51.184801Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-22T12:22:56.972437Z digest=sha256:b39dcef522982846bce0ef5e99bc5957a7996332e9422b58451dc0c884c7ab72

Observation 3b320243-1c8e-436f-8bc9-20baaf6d93fa · inbound

Analytical Provisioning for Attention-FFN Disaggregated LLM Serving under Stochastic Workloads cites this paper.

Analytical Provisioning for Attention-FFN Disaggregated LLM Serving under Stochastic Workloads BurstGPT: A Real-world Workload Dataset to Optimize LLM Serving Systems

Reference 9

Resolution
verified exact
arxiv_id, observed 2026-05-16T09:57:43.041309Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-16T09:53:13.057587Z digest=sha256:501ba52b77fabcbbf013c07520d5c2192bd1f6b88b5f73e8099cffc677e37bc8

Observation 5fb16b47-dbc8-4e6b-9f14-009f7ce6787c · inbound

A Techno-Economic Framework for Cost Modeling and Revenue Opportunities in Open and Programmable AI-RAN cites this paper.

A Techno-Economic Framework for Cost Modeling and Revenue Opportunities in Open and Programmable AI-RAN BurstGPT: A Real-world Workload Dataset to Optimize LLM Serving Systems

Reference 34

Resolution
verified exact
arxiv_id, observed 2026-05-14T00:43:32.076705Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-14T00:41:18.795730Z digest=sha256:5527b113c149cbdceaffa45ac17522847b88488501aec31f564beacd600852b8

Observation 3fd1657e-0628-474d-bcbe-2a55a2842876 · inbound

A Techno-Economic Framework for Cost Modeling and Revenue Opportunities in Open and Programmable AI-RAN cites this paper.

A Techno-Economic Framework for Cost Modeling and Revenue Opportunities in Open and Programmable AI-RAN BurstGPT: A Real-world Workload Dataset to Optimize LLM Serving Systems

Reference 34

Resolution
verified exact
arxiv_id, observed 2026-05-19T16:57:40.185779Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-19T16:54:58.406395Z digest=sha256:f80ff119fcafe7ec20278fe0173d0b25854abba3330873fb567b9bb803ea92d2

Observation 3a3fd922-297f-47cc-b1d0-44dfbbe1ce38 · inbound

Dual-Pool Token-Budget Routing for Cost-Efficient and Reliable LLM Serving cites this paper.

Dual-Pool Token-Budget Routing for Cost-Efficient and Reliable LLM Serving BurstGPT: A Real-world Workload Dataset to Optimize LLM Serving Systems

Reference 26

Resolution
verified exact
arxiv_id, observed 2026-05-11T06:01:02.497582Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-10T17:49:45.604648Z digest=sha256:b0dbfe05965a0ce672c11c1f7bf1953642501e373b68a4a975a1b8b9751e22ee

Observation ee4eb12b-bb62-4539-b8ff-3e76ec3fefab · inbound

When Is the Same Model Not the Same Service? A Measurement Study of Hosted Open-Weight LLM APIs cites this paper.

When Is the Same Model Not the Same Service? A Measurement Study of Hosted Open-Weight LLM APIs BurstGPT: A Real-world Workload Dataset to Optimize LLM Serving Systems

Reference 20

Resolution
verified exact
arxiv_id, observed 2026-05-11T23:11:14.189996Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-08T01:27:54.164991Z digest=sha256:ef2fca5fcfc09bfe4dfba53a23c29ff567d73be39e880b1533bae5834878bdf5

Observation dcf62596-9c8b-42cf-a116-9b9510ef970e · inbound

When Is the Same Model Not the Same Service? A Measurement Study of Hosted Open-Weight LLM APIs cites this paper.

When Is the Same Model Not the Same Service? A Measurement Study of Hosted Open-Weight LLM APIs BurstGPT: A Real-world Workload Dataset to Optimize LLM Serving Systems

Reference 20

Resolution
verified exact
arxiv_id, observed 2026-05-11T17:56:06.140516Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-08T16:57:26.334739Z digest=sha256:fc8597b128e76a7bc9aba7db2ca955ad85dc9a4cce55d304c4847613a323e347

Observation 1adb8f03-79d1-4a81-89f2-4af738e84889 · inbound

ROSE: Rollout On Serving GPUs via Cooperative Elasticity for Agentic RL cites this paper.

ROSE: Rollout On Serving GPUs via Cooperative Elasticity for Agentic RL BurstGPT: A Real-world Workload Dataset to Optimize LLM Serving Systems

Reference 73

Resolution
verified exact
arxiv_id, observed 2026-05-11T21:31:16.212355Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-08T05:14:14.168753Z digest=sha256:61c9dd207a84fd4cc0dd0befd9a281bcc4e585b3fedde7afd24d0335b4ddc373

Observation 34eb5215-4a35-4947-8f3d-1b84e4ed133e · inbound

ROSE: Rollout On Serving GPUs via Cooperative Elasticity for Agentic RL cites this paper.

ROSE: Rollout On Serving GPUs via Cooperative Elasticity for Agentic RL BurstGPT: A Real-world Workload Dataset to Optimize LLM Serving Systems

Reference 72

Resolution
verified exact
arxiv_id, observed 2026-05-21T08:39:53.241704Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-21T08:39:31.911497Z digest=sha256:422c95a069d92ad97db8486a090f3209d71f0d1ba50802fb04a6cfaeaac36f3d

Observation 872c7ed5-4293-4234-8d13-1e34e45fb66b · inbound

ADAPT: A Self-Calibrating Proactive Autoscaler for Container Orchestration cites this paper.

ADAPT: A Self-Calibrating Proactive Autoscaler for Container Orchestration BurstGPT: A Real-world Workload Dataset to Optimize LLM Serving Systems

Reference 10

Resolution
metadata mismatch
arxiv_id, observed 2026-05-19T19:27:43.466513Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-19T19:26:15.876258Z digest=sha256:07a2e0aebcac86c8aa4006ae9ea5f7e27785b791d3a6c92d3a44af1c552cb024

Observation e13b4565-a6f8-4007-b95e-e7e971669804 · inbound

GEM: GPU-Variability-Aware Expert to GPU Mapping for MoE Systems cites this paper.

GEM: GPU-Variability-Aware Expert to GPU Mapping for MoE Systems BurstGPT: A Real-world Workload Dataset to Optimize LLM Serving Systems

Reference 36

Resolution
verified exact
arxiv_id, observed 2026-05-20T04:33:03.850522Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-20T04:31:11.271729Z digest=sha256:1e107982940e45713018cedc384e7e2f7c272d6352a87c08a0622c3f0b0ed44c

Observation 3a858c52-9166-43aa-adb4-c786d7515948 · inbound

Beyond Per-Token Pricing: A Concurrency-Aware Methodology for LLM Infrastructure Cost Estimation cites this paper.

Beyond Per-Token Pricing: A Concurrency-Aware Methodology for LLM Infrastructure Cost Estimation BurstGPT: A Real-world Workload Dataset to Optimize LLM Serving Systems

Reference 34

Resolution
verified exact
arxiv_id, observed 2026-07-03T12:48:11.790687Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-27T08:45:42.781160Z digest=sha256:57d0ab68d26b79af0a7e9af6394b9fa09eaf0c6a028790ddae8d397040263363

Observation bf706967-bcaa-4511-b468-cec43df32685 · inbound

CoCoScale: Leveraging Layer-wise Scaling to Unlock the Potential of Online LLM Serving cites this paper.

CoCoScale: Leveraging Layer-wise Scaling to Unlock the Potential of Online LLM Serving BurstGPT: A Real-world Workload Dataset to Optimize LLM Serving Systems

Reference 31

Resolution
unresolved
no resolver link, observed 2026-07-11T21:08:24.159706Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T21:08:24.159706Z digest=sha256:d105347bbe8a91b8bcb7412efc64523b82a220533eccd21f990c8f9ddb8447be

Observation bed8a48e-ad5b-4099-8241-2d55c718d579 · inbound

Adaptive Inference Batching using Policy Gradients cites this paper.

Adaptive Inference Batching using Policy Gradients BurstGPT: A Real-world Workload Dataset to Optimize LLM Serving Systems

Reference 20

Resolution
metadata mismatch
local_arxiv, observed 2026-07-07T20:34:09.762493Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-07-07T20:29:21.563562Z digest=sha256:b951b27af4b81e52196c0f1a613f60443e82b30fef973662e8ce7870bfed3b29

Observation bcb72876-5e1c-4716-bd4e-1a5261286d4e · inbound

Trusted Floors Under Untrusted Learners: A Runtime Assured-SLO Guard for ML Serving cites this paper.

Trusted Floors Under Untrusted Learners: A Runtime Assured-SLO Guard for ML Serving BurstGPT: A Real-world Workload Dataset to Optimize LLM Serving Systems

Reference 2025

Resolution
unresolved
no resolver link, observed 2026-08-02T07:31:31.140358Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T07:31:31.140358Z digest=sha256:4275c39bc5625cc381e096ff9693e65809dfaa82e78a61bd254860d9259b8d6a

Observation f12090fb-f6f4-4b0c-aa62-55d47fcd2087 · inbound

TokTier: Exact Stateful CPU+GPU Tokenization for Agentic LLM Serving cites this paper.

TokTier: Exact Stateful CPU+GPU Tokenization for Agentic LLM Serving BurstGPT: A Real-world Workload Dataset to Optimize LLM Serving Systems

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-03T02:07:01.398171Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T02:07:01.398171Z digest=sha256:3aecc3c6c375f09df4d15d457bafd9f8492f49009c1edfe99e4b73652f357d85