Pith. sign in

Paper Citation Record · LEDGER

D2O: Dynamic Discriminative Operations for Efficient Long-Context Inference of Large Language Models

As of 10 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 25 inbound Pith citation observations for arXiv:2406.13035.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2406.13035 v3

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 25 of 25 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00

measured 25 of 25 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-10T19:44:35.695799Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

0
arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 018a36cf-76ec-44a7-a8b9-39ad8074989d · inbound

RetroInfer: A Vector Storage Engine for Scalable Long-Context LLM Inference cites this paper.

RetroInfer: A Vector Storage Engine for Scalable Long-Context LLM Inference D2O: Dynamic Discriminative Operations for Efficient Long-Context Inference of Large Language Models

Reference 96

Resolution
verified exact
arxiv_id, observed 2026-05-09T06:35:23.911835Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-22T15:59:04.724780Z digest=sha256:b5854528b6bab5ccbb37ce6467683812e9b3247d181e6c2b852458ec2c88bea4

Observation 735fab92-1f44-4700-825d-a90d83fc65a6 · inbound

Memory-Efficient Visual Autoregressive Modeling with Scale-Aware KV Cache Compression cites this paper.

Memory-Efficient Visual Autoregressive Modeling with Scale-Aware KV Cache Compression D2O: Dynamic Discriminative Operations for Efficient Long-Context Inference of Large Language Models

Reference 57

Resolution
unresolved
no resolver link, observed 2026-08-07T14:16:18.517957Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:16:18.517957Z digest=sha256:072d065492e794222e4a66191518be53e4caf6e17fc56dcf3caa4cb39d6bebf4

Observation 3e7e8cb7-31bb-498a-8848-d0656c204235 · inbound

Cartridges: Lightweight and general-purpose long context representations via self-study cites this paper.

Cartridges: Lightweight and general-purpose long context representations via self-study D2O: Dynamic Discriminative Operations for Efficient Long-Context Inference of Large Language Models

Reference 90

Resolution
unresolved
no resolver link, observed 2026-08-07T06:04:35.044294Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T06:04:35.044294Z digest=sha256:bff96b6773c5c1ff9e49b8f2c8220afbb377e21be04a01a3b6fc7dd4bb8510d0

Observation e7596169-23e7-4742-955b-2e4a21a5c4fa · inbound

SpindleKV: A Novel KV Cache Reduction Method Balancing Both Shallow and Deep Layers cites this paper.

SpindleKV: A Novel KV Cache Reduction Method Balancing Both Shallow and Deep Layers D2O: Dynamic Discriminative Operations for Efficient Long-Context Inference of Large Language Models

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-06T19:10:01.293262Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T19:10:01.293262Z digest=sha256:bfaa3d82fb11ea7985eb57a9892d6bde16e7606a571c209f02ac3ddf744ffdaf

Observation a83163ca-d2e3-4b43-984a-f1c1d963dbcc · inbound

LaCache: Ladder-Shaped KV Caching for Efficient Long-Context Modeling of Large Language Models cites this paper.

LaCache: Ladder-Shaped KV Caching for Efficient Long-Context Modeling of Large Language Models D2O: Dynamic Discriminative Operations for Efficient Long-Context Inference of Large Language Models

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-06T17:33:21.206328Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:33:21.206328Z digest=sha256:9b4cd3938672421095003ce652f1b148951f2b9322accb6a9a5f008077335b85

Observation 0fa505a9-4470-4a51-89ae-1aaee8d2ea88 · inbound

PagedEviction: Structured Block-wise KV Cache Pruning for Efficient Large Language Model Inference cites this paper.

PagedEviction: Structured Block-wise KV Cache Pruning for Efficient Large Language Model Inference D2O: Dynamic Discriminative Operations for Efficient Long-Context Inference of Large Language Models

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-05T10:16:13.103866Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:16:13.103866Z digest=sha256:928b079759fb63e2d2e968b6de90f8af96f97b923b0c292ad433ddbfd7a8714a

Observation bc52c90c-cdb9-465a-b756-f897d130e432 · inbound

STAC: Plug-and-Play Spatio-Temporal Aware Cache Compression for Streaming 3D Reconstruction cites this paper.

STAC: Plug-and-Play Spatio-Temporal Aware Cache Compression for Streaming 3D Reconstruction D2O: Dynamic Discriminative Operations for Efficient Long-Context Inference of Large Language Models

Reference 37

Resolution
verified exact
arxiv_id, observed 2026-05-15T10:05:26.105849Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-15T10:04:47.644492Z digest=sha256:897d559ab293b558ec4c7b8fe35db37339b201257fd1f21f8ac40a62efd9bcb9

Observation 2362e3f5-6144-485f-a0be-5fedfb86efab · inbound

Unifying Sparse Attention with Hierarchical Memory for Scalable Long-Context LLM Serving cites this paper.

Unifying Sparse Attention with Hierarchical Memory for Scalable Long-Context LLM Serving D2O: Dynamic Discriminative Operations for Efficient Long-Context Inference of Large Language Models

Reference 62

Resolution
malformed identifier
arxiv_id, observed 2026-05-12T09:01:25.963272Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-07T13:15:21.201950Z digest=sha256:4a81362e49e15ec77cae04015d8fdff8b588a35b512448b0fe30c9d0db4a3dbf

Observation 25ee5605-c004-4eef-a156-bc2477604c18 · inbound

Reformulating KV Cache Eviction Problem for Long-Context LLM Inference cites this paper.

Reformulating KV Cache Eviction Problem for Long-Context LLM Inference D2O: Dynamic Discriminative Operations for Efficient Long-Context Inference of Large Language Models

Reference 42

Resolution
verified exact
arxiv_id, observed 2026-05-11T03:10:53.553361Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-11T02:37:52.545943Z digest=sha256:5b1f9d6ec7562f270fb6352731fb7d0eee43c5342fae670be0b7e287b3c69c9b

Observation 20fc5fef-40eb-4918-a449-b9e2999d93c3 · inbound

ReST-KV: Robust KV Cache Eviction with Layer-wise Output Reconstruction and Spatial-Temporal Smoothing cites this paper.

ReST-KV: Robust KV Cache Eviction with Layer-wise Output Reconstruction and Spatial-Temporal Smoothing D2O: Dynamic Discriminative Operations for Efficient Long-Context Inference of Large Language Models

Reference 25

Resolution
verified exact
arxiv_id, observed 2026-05-12T07:26:30.402015Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-12T02:52:33.076123Z digest=sha256:a64696937aa06ebdc041ae910e42ac55be55e9bb03b68059d76efc0bf05bfa9d

Observation ab96eddb-6f5a-4ce4-b5a5-3fdfecdd1e04 · inbound

PRISM: Pareto-Efficient Retrieval over Intent-Aware Structured Memory for Long-Horizon Agents cites this paper.

PRISM: Pareto-Efficient Retrieval over Intent-Aware Structured Memory for Long-Horizon Agents D2O: Dynamic Discriminative Operations for Efficient Long-Context Inference of Large Language Models

Reference 22

Resolution
verified exact
arxiv_id, observed 2026-05-13T05:12:17.813605Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-13T05:10:14.521588Z digest=sha256:7a75cef7523a7f78360af104055e4cad40de536c3033a7cbeae2c1311a9decfa

Observation 077fc503-5325-4fd3-93b3-49105a2d0b04 · inbound

PRISM: Pareto-Efficient Retrieval over Intent-Aware Structured Memory for Long-Horizon Agents cites this paper.

PRISM: Pareto-Efficient Retrieval over Intent-Aware Structured Memory for Long-Horizon Agents D2O: Dynamic Discriminative Operations for Efficient Long-Context Inference of Large Language Models

Reference 22

Resolution
verified exact
arxiv_id, observed 2026-05-25T06:25:23.641467Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-25T06:25:07.387150Z digest=sha256:73a87b97702cf753300b89ee0db4203df98cad589bae4fdcc2d89e7b91e30f32

Observation 31e69874-bc1c-4e93-9299-a5dccdf1a889 · inbound

Head-Aware Key-Value Compression for Efficient Autoregressive Image Generation cites this paper.

Head-Aware Key-Value Compression for Efficient Autoregressive Image Generation D2O: Dynamic Discriminative Operations for Efficient Long-Context Inference of Large Language Models

Reference 33

Resolution
verified exact
arxiv_id, observed 2026-05-21T06:09:41.638966Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-21T06:05:23.148656Z digest=sha256:3426baa7a33a22965db8e692302995a929ca0aa34d5efdfbade28013810209c3

Observation 35df0518-695b-45cf-9cca-4e8924192c06 · inbound

Meta-Soft: Leveraging Composable Meta-Tokens for Context-Preserving KV Cache Compression cites this paper.

Meta-Soft: Leveraging Composable Meta-Tokens for Context-Preserving KV Cache Compression D2O: Dynamic Discriminative Operations for Efficient Long-Context Inference of Large Language Models

Reference 24

Resolution
verified exact
arxiv_id, observed 2026-05-22T05:11:06.142848Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-22T05:10:53.133346Z digest=sha256:c326d1d694153e1b942403f2f1434acb9720c944a75e15100c65acea7a5dc00f

Observation 4c1d2cee-4deb-41ad-bbef-c656df2e766f · inbound

Meta-Soft: Leveraging Composable Meta-Tokens for Context-Preserving KV Cache Compression cites this paper.

Meta-Soft: Leveraging Composable Meta-Tokens for Context-Preserving KV Cache Compression D2O: Dynamic Discriminative Operations for Efficient Long-Context Inference of Large Language Models

Reference 24

Resolution
verified exact
arxiv_id, observed 2026-06-30T17:34:57.868384Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-30T17:28:08.750049Z digest=sha256:eb7a91e8f95658a38b3876e542b43269d53d7bd9c5d2a281fa9f3e70c669e9e7

Observation 9e2bea6f-f77d-47be-b5cc-f8b99993e786 · inbound

A Simple Plug-in for Improving Eviction-Based KV Cache Compression cites this paper.

A Simple Plug-in for Improving Eviction-Based KV Cache Compression D2O: Dynamic Discriminative Operations for Efficient Long-Context Inference of Large Language Models

Reference 13

Resolution
verified exact
arxiv_id, observed 2026-05-25T04:55:24.214145Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-25T04:51:04.068354Z digest=sha256:ffbafff8866712b8664711672225a71e619d7faed974584f829d6e0c3dbb7c36

Observation 83cedd67-c583-4531-9e50-cf8b21c2b556 · inbound

LongLive-RAG: A General Retrieval-Augmented Framework for Long Video Generation cites this paper.

LongLive-RAG: A General Retrieval-Augmented Framework for Long Video Generation D2O: Dynamic Discriminative Operations for Efficient Long-Context Inference of Large Language Models

Reference 44

Resolution
verified exact
arxiv_id, observed 2026-07-01T22:16:17.161728Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-28T15:28:41.932618Z digest=sha256:3a61c0566ec68431cfa1c3dd02cc993bc388360c4eb5d289d9062082f1274328

Observation 42c3918b-ca24-4e7b-8c0e-666730244278 · inbound

Language Models Need Sleep: Learning to Self-Modify and Consolidate Memories cites this paper.

Language Models Need Sleep: Learning to Self-Modify and Consolidate Memories D2O: Dynamic Discriminative Operations for Efficient Long-Context Inference of Large Language Models

Reference 96

Resolution
verified exact
arxiv_id, observed 2026-07-02T02:26:26.671308Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-28T10:56:13.058872Z digest=sha256:70a3de0c72d959530dcbd620aa3aba33caf57bb584d4d83d486f6f35d8f884c8

Observation 19fef3ad-0041-42d3-b87f-3566ccf9e60a · inbound

Language Models Need Sleep: Learning to Self-Modify and Consolidate Memories cites this paper.

Language Models Need Sleep: Learning to Self-Modify and Consolidate Memories D2O: Dynamic Discriminative Operations for Efficient Long-Context Inference of Large Language Models

Reference 96

Resolution
unresolved
no resolver link, observed 2026-07-13T07:44:25.325808Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-13T07:44:25.325808Z digest=sha256:79960b26728db63ae19ba9d206f86fffedd6a9f34dba530847a9cf8b61b835a3

Observation f4ec3b23-132e-46f6-9757-5c951222b49b · inbound

HACK++: Towards More Effective Head-Aware Key-Value Compression for Efficient Visual Autoregressive Modeling cites this paper.

HACK++: Towards More Effective Head-Aware Key-Value Compression for Efficient Visual Autoregressive Modeling D2O: Dynamic Discriminative Operations for Efficient Long-Context Inference of Large Language Models

Reference 49

Resolution
verified exact
arxiv_id, observed 2026-07-02T21:27:24.423008Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-27T19:46:43.514413Z digest=sha256:70bf45b8c88e89e4bbaf65a44b78d95e153b351dcd5c1bbb1a9670ff648f5d44

Observation cf9df43f-bb62-4c36-ad74-992ebfd276a3 · inbound

Information-Aware KV Cache Compression for Long Reasoning cites this paper.

Information-Aware KV Cache Compression for Long Reasoning D2O: Dynamic Discriminative Operations for Efficient Long-Context Inference of Large Language Models

Reference 15

Resolution
metadata mismatch
arxiv_id, observed 2026-07-04T13:49:51.574145Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-26T04:55:07.243644Z digest=sha256:861f602cac3af32aa6927eac501378fc3d4ced15e43a93c8e2721802cb1c0fc2

Observation c0ff8af3-6e5d-4a2d-a694-3871d79cd474 · inbound

PhyCheck: Fine-Grained Evidence-Grounded Dataset for Physical Law Understanding in Video-LLMs cites this paper.

PhyCheck: Fine-Grained Evidence-Grounded Dataset for Physical Law Understanding in Video-LLMs D2O: Dynamic Discriminative Operations for Efficient Long-Context Inference of Large Language Models

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-04T13:43:51.754294Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T13:43:51.754294Z digest=sha256:c280081da6d901c5f22e4506e612dd24917213adbab263f3280fb42d16c6414e

Observation 1cca76ca-0a7d-45b1-bf6d-9db3eac69e84 · inbound

PhyCheck: Fine-Grained Evidence-Grounded Dataset for Physical Law Understanding in Video-LLMs cites this paper.

PhyCheck: Fine-Grained Evidence-Grounded Dataset for Physical Law Understanding in Video-LLMs D2O: Dynamic Discriminative Operations for Efficient Long-Context Inference of Large Language Models

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-07T00:11:48.792029Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:11:48.792029Z digest=sha256:e606554d8d845e5f9bbd7292ab72b1942b684425c382b28263a1d7b6dc6107f3

Observation c3d48be3-8fe0-4a3b-af0c-ffec19f407d7 · inbound

Fewer Tokens, Smaller Cache: Reward-Coordinated Efficient Reasoning cites this paper.

Fewer Tokens, Smaller Cache: Reward-Coordinated Efficient Reasoning D2O: Dynamic Discriminative Operations for Efficient Long-Context Inference of Large Language Models

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-06T17:07:01.982112Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:07:01.982112Z digest=sha256:cdabb79a1699a49206caba3e6f261a173db0add6c638bcaf459532a336d10241

Observation 9503d981-62af-4a7d-b25c-49bd3461ac8a · inbound

Autonomy-of-Heads: Data-Free Sparse Attention from Frozen Query-Key Geometry cites this paper.

Autonomy-of-Heads: Data-Free Sparse Attention from Frozen Query-Key Geometry D2O: Dynamic Discriminative Operations for Efficient Long-Context Inference of Large Language Models

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-10T19:44:35.695799Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T19:44:35.695799Z digest=sha256:7fd478e832e69eefbff7260300843102222d546cb9e4e60653ff335b5d9a4e78