Pith. sign in

Paper Citation Record · LEDGER

CITER: Collaborative Inference for Efficient Large Language Model Decoding with Token-Level Routing

As of 6 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 10 inbound Pith citation observations for arXiv:2502.01976.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2502.01976 v6

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 10 of 10 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-06T06:34:29.942622+00:00

measured 10 of 10 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-06T15:06:48.536427Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-01T08:15:32.127896Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation d945a911-9fa0-43ee-bd74-a244d66d22fa · inbound

Harnessing Multiple Large Language Models: A Survey on LLM Ensemble cites this paper.

Harnessing Multiple Large Language Models: A Survey on LLM Ensemble CITER: Collaborative Inference for Efficient Large Language Model Decoding with Token-Level Routing

Reference 70

Resolution
verified exact
arxiv_id, observed 2026-05-23T02:25:19.669400Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-23T02:22:28.649071Z digest=sha256:e2368844679bd52c22f46fff964abdc66714cf5690218c661ca38d866702594e

Observation bb73bfbc-065e-45c1-8575-0053ad623da8 · inbound

Sampling from Your Language Model One Byte at a Time cites this paper.

Sampling from Your Language Model One Byte at a Time CITER: Collaborative Inference for Efficient Large Language Model Decoding with Token-Level Routing

Reference 84

Resolution
verified exact
arxiv_id, observed 2026-05-19T09:53:02.078562Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-19T09:52:20.923710Z digest=sha256:b7061996aac81b9aa7f86abfb004dcbbaa3a836292feba89310d61d3864b79c5

Observation 127465ed-f961-4ed3-8f79-c835b9b7616f · inbound

Collaborative Inference and Learning between Edge SLMs and Cloud LLMs: A Survey of Algorithms, Execution, and Open Challenges cites this paper.

Collaborative Inference and Learning between Edge SLMs and Cloud LLMs: A Survey of Algorithms, Execution, and Open Challenges CITER: Collaborative Inference for Efficient Large Language Model Decoding with Token-Level Routing

Reference 183

Resolution
unresolved
no resolver link, observed 2026-08-06T15:06:48.536427Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:06:48.536427Z digest=sha256:331652431d69cc862a358006c51ded2b0b69e705b487cef4294bf337914b92aa

Observation 2df8c9ca-e25c-4918-8ad8-efbfd166a930 · inbound

NI Sampling: Accelerating Discrete Diffusion Sampling by Token Order Optimization cites this paper.

NI Sampling: Accelerating Discrete Diffusion Sampling by Token Order Optimization CITER: Collaborative Inference for Efficient Large Language Model Decoding with Token-Level Routing

Reference 33

Resolution
metadata mismatch
arxiv_id, observed 2026-05-10T06:01:13.633017Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-05-10T05:56:52.205196Z digest=sha256:a29d40739439c7ef11b206bd53b2096f1aca4b12a7a77646a0b20b773375760d

Observation 9482e410-3884-4ffd-b979-89f5c33f1c2a · inbound

Rethinking LLM Ensembling from the Perspective of Mixture Models cites this paper.

Rethinking LLM Ensembling from the Perspective of Mixture Models CITER: Collaborative Inference for Efficient Large Language Model Decoding with Token-Level Routing

Reference 18

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T15:26:06.339051Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-09T20:06:12.248439Z digest=sha256:8608bc31f8e3b871e4cf1f11fbeeb219c011c4187c79e58bcacd5a0808e56056

Observation 1e596ca0-c9d6-4902-8889-de3a477d5455 · inbound

Rethinking LLM Ensembling from the Perspective of Mixture Models cites this paper.

Rethinking LLM Ensembling from the Perspective of Mixture Models CITER: Collaborative Inference for Efficient Large Language Model Decoding with Token-Level Routing

Reference 18

Resolution
metadata mismatch
arxiv_id, observed 2026-07-01T08:15:32.130199Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-07-01T08:07:09.032028Z digest=sha256:6a51e476e10ae33496f359f200d07dc030aed34840e7dd77d589c593ef7577a1

Observation 28555106-1416-4f58-b7f3-19a2b0208d0c · inbound

Accelerating Heterogeneous Agent Collaboration in Dynamic Edge Networks cites this paper.

Accelerating Heterogeneous Agent Collaboration in Dynamic Edge Networks CITER: Collaborative Inference for Efficient Large Language Model Decoding with Token-Level Routing

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-02T15:18:54.607262Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T15:18:54.607262Z digest=sha256:1b8c267413c2dde31723817c37899b7ecf4f34a17a6162269417d955c612c73e

Observation 9791e570-20ef-4b91-9855-395ec8fb6379 · inbound

PyroDash: Cost-Efficient Token-Level Small-Large Language Model Collaborative Inference cites this paper.

PyroDash: Cost-Efficient Token-Level Small-Large Language Model Collaborative Inference CITER: Collaborative Inference for Efficient Large Language Model Decoding with Token-Level Routing

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-01T10:14:15.660311Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T10:14:15.660311Z digest=sha256:144beeaf96727d9e660c1170767c207e37c1e357e30ce59e52df60cad9d34796

Observation 07fd0cd4-105c-4dbf-bb4d-824d58d5dc52 · inbound

TRACE-ROUTER: Task-Consistent and Adaptive Online Routing for Agentic AI cites this paper.

TRACE-ROUTER: Task-Consistent and Adaptive Online Routing for Agentic AI CITER: Collaborative Inference for Efficient Large Language Model Decoding with Token-Level Routing

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-01T04:42:15.122849Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T04:42:15.122849Z digest=sha256:ed9aef044dac00f9f36dc9569c60eb4276d1baa3f32687a0a3b82a5719a321ae

Observation c2049d7d-e5e1-44e9-9f02-3fe40069efe6 · inbound

Divergence Decoding: Training-Free Capability Fusion cites this paper.

Divergence Decoding: Training-Free Capability Fusion CITER: Collaborative Inference for Efficient Large Language Model Decoding with Token-Level Routing

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-01T02:47:11.302112Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T02:47:11.302112Z digest=sha256:aaa86a49b153bcc1569ca57d119d0c2b3d27568d39c37307c09ac186a3e8035a