Pith. sign in

Paper Citation Record · LEDGER

Discovering the Gems in Early Layers: Accelerating Long-Context LLMs with 1000x Input Token Reduction

As of 7 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 9 inbound Pith citation observations for arXiv:2409.17422.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2409.17422 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 9 of 9 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 9 of 9 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T12:39:05.601985Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

0
arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 7f7592e2-bbfe-468e-8737-9d5c95431b5b · inbound

FastKV: Decoupling of Context Reduction and KV Cache Compression for Prefill-Decoding Acceleration cites this paper.

FastKV: Decoupling of Context Reduction and KV Cache Compression for Prefill-Decoding Acceleration Discovering the Gems in Early Layers: Accelerating Long-Context LLMs with 1000x Input Token Reduction

Reference 29

Resolution
verified exact
arxiv_id, observed 2026-05-23T04:02:30.316490Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-05-23T03:59:58.634512Z digest=sha256:dcf6e0342ac2f955905c9ff61893af173b1d9cb23fb56ea40847c35122587977

Observation 0e17897b-9f3b-4af4-bd78-0b15be0fe53d · inbound

SALE : Low-bit Estimation for Efficient Sparse Attention in Long-context LLM Prefilling cites this paper.

SALE : Low-bit Estimation for Efficient Sparse Attention in Long-context LLM Prefilling Discovering the Gems in Early Layers: Accelerating Long-Context LLMs with 1000x Input Token Reduction

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-07T12:39:05.601985Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:39:05.601985Z digest=sha256:73f8b032304faeffe4e28b02dbe86bbf71236f0f1b292aaf195302a15a2b310c

Observation 88163e8c-8b7b-431e-85a9-a0678b69b679 · inbound

EARN: Efficient Inference Acceleration for LLM-based Generative Recommendation by Register Tokens cites this paper.

EARN: Efficient Inference Acceleration for LLM-based Generative Recommendation by Register Tokens Discovering the Gems in Early Layers: Accelerating Long-Context LLMs with 1000x Input Token Reduction

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-06T21:18:50.235479Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:18:50.235479Z digest=sha256:7257dd7f6bbdbe326b3db2e000f61faf9412bbb9a285f9bc4c1532160724231a

Observation 0de5561e-521a-48b1-9f0f-5b1ec5ea8bb5 · inbound

Unifying Learning Dynamics and Generalization in Transformers Scaling Law cites this paper.

Unifying Learning Dynamics and Generalization in Transformers Scaling Law Discovering the Gems in Early Layers: Accelerating Long-Context LLMs with 1000x Input Token Reduction

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-03T14:02:56.724087Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T14:02:56.724087Z digest=sha256:febff48b15c052ab71536972cff079fd65d596f4f96c596f0c74a0ffb3aa84a4

Observation d7b80cdd-66bb-46ee-abcd-7fb5b0da0ecd · inbound

Token Sparse Attention: Efficient Long-Context Inference with Interleaved Token Selection cites this paper.

Token Sparse Attention: Efficient Long-Context Inference with Interleaved Token Selection Discovering the Gems in Early Layers: Accelerating Long-Context LLMs with 1000x Input Token Reduction

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-03T05:09:31.306514Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-03T05:09:31.306514Z digest=sha256:8e9775a31df60b1de62ec144ada391be2c01745679e792d1b64dd7e5142b2bc7

Observation 5fab0228-081c-4279-93da-8661da3c3267 · inbound

StructKV: Preserving the Structural Skeleton for Scalable Long-Context Inference cites this paper.

StructKV: Preserving the Structural Skeleton for Scalable Long-Context Inference Discovering the Gems in Early Layers: Accelerating Long-Context LLMs with 1000x Input Token Reduction

Reference 15

Resolution
verified exact
arxiv_id, observed 2026-05-10T18:40:44.112794Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-05-10T18:37:10.526359Z digest=sha256:316382c7579c91acb76e49c101456b0b5d303d995fead77f490fbc50b08550a8

Observation a5846ad2-7d7f-4094-a949-38706c878d54 · inbound

Correctness-Aware Repository Filtering Under Maximum Effective Context Window Constraints cites this paper.

Correctness-Aware Repository Filtering Under Maximum Effective Context Window Constraints Discovering the Gems in Early Layers: Accelerating Long-Context LLMs with 1000x Input Token Reduction

Reference 19

Resolution
verified exact
arxiv_id, observed 2026-05-15T02:38:34.603206Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-15T02:37:57.481847Z digest=sha256:586971dfa3a187b2de7d8763e796981af0cc1c4b3b1f5dea66b11bb686da11fe

Observation 012e4a30-cb49-4d45-8aa3-7b1cf9a6d113 · inbound

TokenMizer: Graph-Structured Session Memory for Long-Horizon LLM Context Management cites this paper.

TokenMizer: Graph-Structured Session Memory for Long-Horizon LLM Context Management Discovering the Gems in Early Layers: Accelerating Long-Context LLMs with 1000x Input Token Reduction

Reference 20

Resolution
verified exact
arxiv_id, observed 2026-07-02T13:06:59.133394Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-28T01:35:01.565343Z digest=sha256:935251d7ccc35a9b7bd6ef4cc0b0912869c18212f8cf78c4d0e9fa02b12e2925

Observation 0828e6ad-9792-486b-b635-f8eed2672a5c · inbound

Coverage-Driven KV Cache Eviction for Efficient and Improved Inference of LLM cites this paper.

Coverage-Driven KV Cache Eviction for Efficient and Improved Inference of LLM Discovering the Gems in Early Layers: Accelerating Long-Context LLMs with 1000x Input Token Reduction

Reference 78

Resolution
verified exact
arxiv_id, observed 2026-06-30T07:24:21.357324Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-06-30T07:19:48.530272Z digest=sha256:ff1d3c5661e4f993dfef73fb220336e8dbd5c2e736dbcae18fe8c18c84a1463a