Pith. sign in

Paper Citation Record · LEDGER

InferCept: Efficient Intercept Support for Augmented Large Language Model Inference

As of 7 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 9 inbound Pith citation observations for arXiv:2402.01869.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2402.01869 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 9 of 9 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-06T06:34:29.942622+00:00

measured 9 of 9 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-06T18:45:42.384225Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-02T17:27:14.649493Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 6a74ddd4-1c8b-4d04-85ea-4a396ab5641c · inbound

SGLang: Efficient Execution of Structured Language Model Programs cites this paper.

SGLang: Efficient Execution of Structured Language Model Programs InferCept: Efficient Intercept Support for Augmented Large Language Model Inference

Reference 1

Resolution
verified exact
arxiv_id, observed 2026-05-12T08:20:01.189749Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-12T08:20:01.011625Z digest=sha256:ef2b7bc7b7a2d36477f2bb7d9b1a7c294ef3c7d4541c0f2ebf7ed62ab509d537

Observation 519ae05f-2cce-46bd-92ab-866c62479560 · inbound

KVFlow: Efficient Prefix Caching for Accelerating LLM-Based Multi-Agent Workflows cites this paper.

KVFlow: Efficient Prefix Caching for Accelerating LLM-Based Multi-Agent Workflows InferCept: Efficient Intercept Support for Augmented Large Language Model Inference

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-06T18:45:42.384225Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:45:42.384225Z digest=sha256:ba875c131e72f8651240d119c531ab3af5c4a0cbf42341548fe58e521be4ab2e

Observation ac1324be-e345-4427-b7bb-471952c47785 · inbound

AgentOpt v0.1 Technical Report: Client-Side Optimization for LLM-Based Agent cites this paper.

AgentOpt v0.1 Technical Report: Client-Side Optimization for LLM-Based Agent InferCept: Efficient Intercept Support for Augmented Large Language Model Inference

Reference 1

Resolution
verified exact
arxiv_id, observed 2026-05-10T22:25:52.677083Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T19:51:39.564680Z digest=sha256:85bc77643231820c4046913a29d7cc798e2bce09105c4a2fb7e619e168bccdb2

Observation ac779d41-064d-40c0-9e94-9f85fa18e7bf · inbound

Diffusion Templates: A Unified Plugin Framework for Controllable Diffusion cites this paper.

Diffusion Templates: A Unified Plugin Framework for Controllable Diffusion InferCept: Efficient Intercept Support for Augmented Large Language Model Inference

Reference 1

Resolution
verified exact
arxiv_id, observed 2026-05-11T21:46:42.166113Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-08T04:24:53.023047Z digest=sha256:716a2a19f4e35a5f4303b02ac71b0cc84d2722af63b3540b18b3831699d7ad38

Observation 46930153-0e49-4a2c-9564-5609816785d0 · inbound

A Policy-Driven Runtime Layer for Agentic LLM Serving cites this paper.

A Policy-Driven Runtime Layer for Agentic LLM Serving InferCept: Efficient Intercept Support for Augmented Large Language Model Inference

Reference 1

Resolution
verified exact
arxiv_id, observed 2026-06-29T17:03:41.382323Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-06-29T16:54:36.387432Z digest=sha256:1ce4c1d150af9dcb5f2394ce644006ae880312156fcb890c3bb055c49d69153b

Observation 204b59b4-e6e0-4ecb-8051-4d428743ce86 · inbound

A Policy-Driven Runtime Layer for Agentic LLM Serving cites this paper.

A Policy-Driven Runtime Layer for Agentic LLM Serving InferCept: Efficient Intercept Support for Augmented Large Language Model Inference

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-02T13:00:09.691002Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T13:00:09.691002Z digest=sha256:54e44e29cbb47f91fd19f179159ff86e0c46a5bb17e73c24f57959975afe22ab

Observation 4d622906-dd7a-4d13-92a5-c82ef518ac03 · inbound

Idleness is Relative: Exploiting Tool-Call Idle Windows for Offloading in Agentic Systems with MORI cites this paper.

Idleness is Relative: Exploiting Tool-Call Idle Windows for Offloading in Agentic Systems with MORI InferCept: Efficient Intercept Support for Augmented Large Language Model Inference

Reference 1

Resolution
verified exact
arxiv_id, observed 2026-07-01T21:06:13.642703Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-06-28T17:30:56.324289Z digest=sha256:34c9d3f03bbcbd8207395fe8732d07031b738559368ed3a95c68ee07a593ad0c

Observation 56d05b36-8b88-4215-a6f4-2601c2a33f3e · inbound

Multi-Segment Attention: Enabling Efficient KV-Cache Management for Faster Large Language Model Serving cites this paper.

Multi-Segment Attention: Enabling Efficient KV-Cache Management for Faster Large Language Model Serving InferCept: Efficient Intercept Support for Augmented Large Language Model Inference

Reference 4

Resolution
verified exact
arxiv_id, observed 2026-07-02T01:36:26.274881Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-06-28T11:38:05.435505Z digest=sha256:b5faac8959babb638832ae58cfec620b466edb5c95e171f10ab6cfe2a08cd772

Observation 5c8e7933-6694-4643-9fc8-0a52ec073e23 · inbound

SmoothAgent: Efficient Long-Horizon LLM-Based Agent Serving with Lookahead Context Engineering cites this paper.

SmoothAgent: Efficient Long-Horizon LLM-Based Agent Serving with Lookahead Context Engineering InferCept: Efficient Intercept Support for Augmented Large Language Model Inference

Reference 2

Resolution
verified exact
arxiv_id, observed 2026-07-02T17:27:14.652081Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-07-02T17:19:53.560039Z digest=sha256:efa96c490e1e1750c70b2509007ae901d77d14e4a3732be27728d7bc75df0bc9