Pith. sign in

Paper Citation Record · LEDGER

Hydragen: High-Throughput LLM Inference with Shared Prefixes

As of 4 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 13 inbound Pith citation observations for arXiv:2402.05099.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2402.05099 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 13 of 13 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-04T06:34:03.388597+00:00

measured 13 of 13 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-07-12T09:50:23.266920Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-04T11:39:46.505898Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 8dc0df40-83bd-4002-b1a2-e7021f3f8de3 · inbound

SGLang: Efficient Execution of Structured Language Model Programs cites this paper.

SGLang: Efficient Execution of Structured Language Model Programs Hydragen: High-Throughput LLM Inference with Shared Prefixes

Reference 18

Resolution
metadata mismatch
arxiv_id, observed 2026-05-12T08:20:01.176768Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-12T08:20:01.011625Z digest=sha256:9e574e5d8125907bc46edd740b8b251a56a96f18883c8503a7fa1ca03820fefc

Observation f7378c09-19ef-4bdd-99d8-7d21e8faacae · inbound

Large Language Monkeys: Scaling Inference Compute with Repeated Sampling cites this paper.

Large Language Monkeys: Scaling Inference Compute with Repeated Sampling Hydragen: High-Throughput LLM Inference with Shared Prefixes

Reference 37

Resolution
metadata mismatch
arxiv_id, observed 2026-05-12T04:42:23.616982Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-12T04:42:23.297389Z digest=sha256:311dc6632d857d4a7864ca9b43e4ecc1815ef011dd3baad9d65955c33a098eb0

Observation 29ad9e26-4fd4-4c5b-bb39-9adf2a7d1edd · inbound

BatchLLM: Optimizing Large Batched LLM Inference with Global Prefix Sharing and Throughput-oriented Token Batching cites this paper.

BatchLLM: Optimizing Large Batched LLM Inference with Global Prefix Sharing and Throughput-oriented Token Batching Hydragen: High-Throughput LLM Inference with Shared Prefixes

Reference 16

Resolution
verified exact
arxiv_id, observed 2026-05-23T16:58:11.956282Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-23T16:57:46.645061Z digest=sha256:47cd823285f8f298d1a5f9c8be71f0dd4860acd27e8be95e002c47025cb13a84

Observation b29408a1-2498-4e35-9655-7eebfbf5dd8d · inbound

PrefixWall: Mitigating Prefix Caching Side Channels in Shared LLM Systems cites this paper.

PrefixWall: Mitigating Prefix Caching Side Channels in Shared LLM Systems Hydragen: High-Throughput LLM Inference with Shared Prefixes

Reference 38

Resolution
verified exact
arxiv_id, observed 2026-05-21T12:15:06.784846Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-21T12:14:09.509302Z digest=sha256:f1706d8e2dda7f0457465d051aa5a76488ce45155dee3dc9b1e76de519459aa7

Observation 871fe2b3-6d42-4b7d-9d84-4e326f4e8d49 · inbound

SEMA-SQL: Beyond Traditional Relational Querying with Large Language Models cites this paper.

SEMA-SQL: Beyond Traditional Relational Querying with Large Language Models Hydragen: High-Throughput LLM Inference with Shared Prefixes

Reference 23

Resolution
verified exact
arxiv_id, observed 2026-05-11T21:31:17.460759Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-08T05:11:46.991452Z digest=sha256:6e8098231ed698cefaf4809091d6a9deac5d9aac0f9ab69407727d78061988fe

Observation 215da78f-f1e6-4f31-bf4d-dd1394a001fc · inbound

SEMA-SQL: Beyond Traditional Relational Querying with Large Language Models cites this paper.

SEMA-SQL: Beyond Traditional Relational Querying with Large Language Models Hydragen: High-Throughput LLM Inference with Shared Prefixes

Reference 23

Resolution
verified exact
arxiv_id, observed 2026-05-15T06:55:10.559212Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-15T06:53:11.604043Z digest=sha256:658f3c7a1115908284e3ee571b5741869bbb3e3f20148abdd501b5f63109df29

Observation cdedbf31-6a5a-4d3f-bd19-7c5b09b0f9e7 · inbound

MoE-Prefill: Zero Redundancy Overheads in MoE Prefill Serving cites this paper.

MoE-Prefill: Zero Redundancy Overheads in MoE Prefill Serving Hydragen: High-Throughput LLM Inference with Shared Prefixes

Reference 32

Resolution
metadata mismatch
arxiv_id, observed 2026-05-09T05:50:27.025286Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-08T19:29:28.916831Z digest=sha256:9b4ffa67fb1118aad16174893cf91730dfef8a0201083fffeb54edc2560eb932

Observation bf46b8a7-09f8-4d97-b749-180b2cc8fe59 · inbound

MoE-Prefill: Zero Redundancy Overheads in MoE Prefill Serving cites this paper.

MoE-Prefill: Zero Redundancy Overheads in MoE Prefill Serving Hydragen: High-Throughput LLM Inference with Shared Prefixes

Reference 32

Resolution
metadata mismatch
arxiv_id, observed 2026-05-19T17:47:41.948204Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-19T17:43:56.767714Z digest=sha256:8188828576384eb81bba90a3ea97d1a6d799dd9684ce8e32c0a37de328ea7460

Observation f1d616d3-9ae6-4017-9f98-73d82614dd82 · inbound

Towards Distributed Inference of LLMs on a P2P Network cites this paper.

Towards Distributed Inference of LLMs on a P2P Network Hydragen: High-Throughput LLM Inference with Shared Prefixes

Reference 5

Resolution
metadata mismatch
arxiv_id, observed 2026-06-30T23:15:07.948549Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-06-30T23:13:57.705287Z digest=sha256:3702c66d223ac03673fa6ec79129c3838e388d66af98242f62f795e617050534

Observation 2da98e4e-03c3-407d-82ff-7497334a5223 · inbound

Execution-State Capsules: Graph-Bound Execution-State Checkpoint and Restore for Low-Latency, Small-Batch, On-Device Physical-AI Serving cites this paper.

Execution-State Capsules: Graph-Bound Execution-State Checkpoint and Restore for Low-Latency, Small-Batch, On-Device Physical-AI Serving Hydragen: High-Throughput LLM Inference with Shared Prefixes

Reference 5

Resolution
verified exact
arxiv_id, observed 2026-07-04T03:49:29.915609Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-06-26T17:39:47.477338Z digest=sha256:130000a3eed205baeac1d0452d4ef5dcc9200e434965dbcc754b3f5595f8bcdc

Observation 1f20be0c-1054-4a50-b7b3-0d28561a3184 · inbound

One Generator, Any Process: LLM-Conditioning for the LHC cites this paper.

One Generator, Any Process: LLM-Conditioning for the LHC Hydragen: High-Throughput LLM Inference with Shared Prefixes

Reference 276

Resolution
verified exact
arxiv_id, observed 2026-07-04T11:39:46.507262Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-06-26T07:53:57.250401Z digest=sha256:cfe703eac44d219d1bca2ae366a67aeec75bd2df6838a995bf5bbe57ce44c853

Observation 1157fc1f-5e5f-40d0-85bc-7aee4a5b97ca · inbound

One Generator, Any Process: LLM-Conditioning for the LHC cites this paper.

One Generator, Any Process: LLM-Conditioning for the LHC Hydragen: High-Throughput LLM Inference with Shared Prefixes

Reference 280

Resolution
verified exact
arxiv_id, observed 2026-06-30T10:14:36.110959Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-06-30T10:13:09.503522Z digest=sha256:8ef1a08d32358b076f19cbb08c762f8a06ba54ae2c2b91cbb0e6a1dac419b1ba

Observation ec8e962e-f30c-442b-9bf0-228e2aebb6a8 · inbound

From Tensor Buffer to Distributed Memory Hierarchy: A Survey of KV Cache Management for LLM Serving cites this paper.

From Tensor Buffer to Distributed Memory Hierarchy: A Survey of KV Cache Management for LLM Serving Hydragen: High-Throughput LLM Inference with Shared Prefixes

Reference 55

Resolution
unresolved
no resolver link, observed 2026-07-12T09:50:23.266920Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T09:50:23.266920Z digest=sha256:5cb897c4da2cff156f0e3e170d3461bf65a5dd086febd6f49f5aafab7244bb03