Pith. sign in

Paper Citation Record · LEDGER

LOOK-M: Look-Once Optimization in KV Cache for Efficient Multimodal Long-Context Inference

As of 7 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 21 inbound Pith citation observations for arXiv:2406.18139.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2406.18139 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 21 of 21 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 21 of 21 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T15:15:11.922497Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-07-10T01:36:44.109589Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 97fa194d-5d25-4a7e-8d84-69069a3bbfd9 · inbound

When Attention Sink Emerges in Language Models: An Empirical View cites this paper.

When Attention Sink Emerges in Language Models: An Empirical View LOOK-M: Look-Once Optimization in KV Cache for Efficient Multimodal Long-Context Inference

Reference 50

Resolution
verified exact
arxiv_id, observed 2026-05-16T17:41:03.758200Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-05-16T17:41:03.674759Z digest=sha256:1a678b0f6282bceaeaf697b72c9ddfd156ef35a9aee5adaca3b8d9e24704b0fe

Observation e0eeda1f-9f52-4ae6-9519-2ecb36638b25 · inbound

Streamline Without Sacrifice -- Squeeze out Computation Redundancy in LMM cites this paper.

Streamline Without Sacrifice -- Squeeze out Computation Redundancy in LMM LOOK-M: Look-Once Optimization in KV Cache for Efficient Multimodal Long-Context Inference

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-07T15:15:11.922497Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:15:11.922497Z digest=sha256:718d091ad4c7da909c8af96bc437a91d5a0bb772944638a1f3f41c91270b4c7f

Observation 1cabc8b2-404d-4115-8940-3fe58a98ed21 · inbound

Memory-Efficient Visual Autoregressive Modeling with Scale-Aware KV Cache Compression cites this paper.

Memory-Efficient Visual Autoregressive Modeling with Scale-Aware KV Cache Compression LOOK-M: Look-Once Optimization in KV Cache for Efficient Multimodal Long-Context Inference

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-07T14:16:18.606283Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:16:18.606283Z digest=sha256:d10c643d86c0e8829790a19f26996041da68c58fe6c3982381e62e3bd5158b9e

Observation eaf5978a-e6b8-4f6c-928f-420226815c55 · inbound

Zooming from Context to Cue: Hierarchical Preference Optimization for Multi-Image MLLMs cites this paper.

Zooming from Context to Cue: Hierarchical Preference Optimization for Multi-Image MLLMs LOOK-M: Look-Once Optimization in KV Cache for Efficient Multimodal Long-Context Inference

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-07T13:14:05.941977Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:14:05.941977Z digest=sha256:5118460fb24cdbd5b34e50c6f973d66242ecd359aaff4d93630310ae38221b04

Observation b3ee3db6-3c52-4b84-865b-109953407c42 · inbound

MadaKV: Adaptive Modality-Perception KV Cache Eviction for Efficient Multimodal Long-Context Inference cites this paper.

MadaKV: Adaptive Modality-Perception KV Cache Eviction for Efficient Multimodal Long-Context Inference LOOK-M: Look-Once Optimization in KV Cache for Efficient Multimodal Long-Context Inference

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-07T10:20:43.864806Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:20:43.864806Z digest=sha256:34922fcc78d5d49074b0b0599b7a1f5148094eb2e1e9fd096c2a93b568d34786

Observation fdea2b54-0b7f-49e1-965a-c3998b9c5e04 · inbound

LaVi: Efficient Large Vision-Language Models via Internal Feature Modulation cites this paper.

LaVi: Efficient Large Vision-Language Models via Internal Feature Modulation LOOK-M: Look-Once Optimization in KV Cache for Efficient Multimodal Long-Context Inference

Reference 61

Resolution
unresolved
no resolver link, observed 2026-08-06T23:42:13.894002Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:42:13.894002Z digest=sha256:08d19587501f3d0f6e1cae3ef85f0b25f7e5d290f9acf1f631041e948c75d291

Observation b265b223-eb3a-43b3-b1eb-2b0e57b0a087 · inbound

Breaking the Boundaries of Long-Context LLM Inference: Adaptive KV Management on a Single Commodity GPU cites this paper.

Breaking the Boundaries of Long-Context LLM Inference: Adaptive KV Management on a Single Commodity GPU LOOK-M: Look-Once Optimization in KV Cache for Efficient Multimodal Long-Context Inference

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-06T23:00:11.193142Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:00:11.193142Z digest=sha256:ce104b75f3ffc2063f96a7f7672524b59c572355cabd56cad32124d107cb3789

Observation 50d62950-433e-436a-9662-4ce9c855f577 · inbound

Multi-Granular Spatio-Temporal Token Merging for Training-Free Acceleration of Video LLMs cites this paper.

Multi-Granular Spatio-Temporal Token Merging for Training-Free Acceleration of Video LLMs LOOK-M: Look-Once Optimization in KV Cache for Efficient Multimodal Long-Context Inference

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-06T18:32:49.280218Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:32:49.280218Z digest=sha256:0dfc0009d872fab559d3056b4d2a69be5f36bb1b4e830d705c812b0ef0b61754

Observation 288fe766-d424-4dab-bd1f-25d092fcba27 · inbound

POINTS-Long: Adaptive Dual-Mode Visual Reasoning in MLLMs cites this paper.

POINTS-Long: Adaptive Dual-Mode Visual Reasoning in MLLMs LOOK-M: Look-Once Optimization in KV Cache for Efficient Multimodal Long-Context Inference

Reference 83

Resolution
verified exact
arxiv_id, observed 2026-05-11T10:41:03.814093Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-10T15:23:08.671342Z digest=sha256:8ba689a73b3fa74181ed7addec11d24e5b0799807f9fcc301474c82dccf24b4f

Observation 21c60f4b-2a68-41c3-ab1a-f1f0fae2b42c · inbound

Reducing Peak Memory Usage for Modern Multimodal Large Language Model Pipelines cites this paper.

Reducing Peak Memory Usage for Modern Multimodal Large Language Model Pipelines LOOK-M: Look-Once Optimization in KV Cache for Efficient Multimodal Long-Context Inference

Reference 3

Resolution
metadata mismatch
arxiv_id, observed 2026-05-10T08:17:37.412398Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-10T08:16:02.102975Z digest=sha256:e882e182acb758a65a31b59e563481a4ea7315301152b3a857fa4c4d16db36a7

Observation 877a345f-8fb9-460f-86e1-7b63debec640 · inbound

Geometry-Guided 3D Visual Token Pruning for Video-Language Models cites this paper.

Geometry-Guided 3D Visual Token Pruning for Video-Language Models LOOK-M: Look-Once Optimization in KV Cache for Efficient Multimodal Long-Context Inference

Reference 27

Resolution
verified exact
arxiv_id, observed 2026-05-10T05:51:09.954656Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-10T05:49:38.346274Z digest=sha256:e83470293ff97ed970425f987396a228297dbb980393d89db5bbafd3464145fd

Observation 8fc9dade-a558-48f2-aa43-ee39443b6fcb · inbound

Prefill-Time Intervention for Mitigating Hallucination in Large Vision-Language Models cites this paper.

Prefill-Time Intervention for Mitigating Hallucination in Large Vision-Language Models LOOK-M: Look-Once Optimization in KV Cache for Efficient Multimodal Long-Context Inference

Reference 44

Resolution
verified exact
arxiv_id, observed 2026-05-11T23:31:16.285028Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-07T16:50:26.044591Z digest=sha256:6d97787f31ba848da90cf05cf1b989660d64397a2743e98a499dbf4cda658b90

Observation b76fe186-f0f4-48b4-8d16-e9c3f2ee60b0 · inbound

RetentiveKV: State-Space Memory for Uncertainty-Aware Multimodal KV Cache Eviction cites this paper.

RetentiveKV: State-Space Memory for Uncertainty-Aware Multimodal KV Cache Eviction LOOK-M: Look-Once Optimization in KV Cache for Efficient Multimodal Long-Context Inference

Reference 51

Resolution
verified exact
arxiv_id, observed 2026-05-11T10:26:02.116725Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-05-10T15:29:17.567557Z digest=sha256:c95fd537973ca4076e361ee2c4d55d694e9de836c8dc64c2f7c77fb4e16d0ef7

Observation 83b6d075-edef-44c0-aa23-0f1e797cb643 · inbound

The Structural Origin of Attention Sink: Variance Discrepancy, Super Neurons, and Dimension Disparity cites this paper.

The Structural Origin of Attention Sink: Variance Discrepancy, Super Neurons, and Dimension Disparity LOOK-M: Look-Once Optimization in KV Cache for Efficient Multimodal Long-Context Inference

Reference 27

Resolution
verified exact
arxiv_id, observed 2026-05-11T19:21:08.583429Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-08T12:11:04.146711Z digest=sha256:bbbe36d2503bc880502e35777a560ff53688f23fdcb51f10c856a42dd4a1b7b0

Observation 420c0c10-6f9d-40d2-b6af-f9b97f0c0fc8 · inbound

KVCapsule: Efficient Sequential KV Cache Compression for Vision-Language Models with Asymmetric Redundancy cites this paper.

KVCapsule: Efficient Sequential KV Cache Compression for Vision-Language Models with Asymmetric Redundancy LOOK-M: Look-Once Optimization in KV Cache for Efficient Multimodal Long-Context Inference

Reference 28

Resolution
verified exact
arxiv_id, observed 2026-05-20T19:58:59.074345Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-20T19:55:35.051834Z digest=sha256:3c7fd1ba0d426380c24a862d7e8a5a97b18dc0274c11ddf3c6f8f5ffe45a7907

Observation 1c2908e8-a975-435b-a5e8-182033b85bd2 · inbound

VisionPulse: Dynamic Visual Sparsity for Efficient Multimodal Reasoning cites this paper.

VisionPulse: Dynamic Visual Sparsity for Efficient Multimodal Reasoning LOOK-M: Look-Once Optimization in KV Cache for Efficient Multimodal Long-Context Inference

Reference 11

Resolution
verified exact
arxiv_id, observed 2026-06-29T00:12:50.575478Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-28T23:13:15.593994Z digest=sha256:ca29e7c1846ef83a331a84e739b2a8df4a8abceb26b67bc443900b2cb9d3b9ec

Observation 3c964e99-c6c0-42ff-9c93-a4a254a53dc9 · inbound

HACK++: Towards More Effective Head-Aware Key-Value Compression for Efficient Visual Autoregressive Modeling cites this paper.

HACK++: Towards More Effective Head-Aware Key-Value Compression for Efficient Visual Autoregressive Modeling LOOK-M: Look-Once Optimization in KV Cache for Efficient Multimodal Long-Context Inference

Reference 46

Resolution
verified exact
arxiv_id, observed 2026-07-02T21:27:24.429651Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-27T19:46:43.514413Z digest=sha256:2d49ef96611e13aced6470bfb5b1a0eaa0bda36f0c96324cd6527a123603eb3d

Observation bd0d4249-644e-4203-951f-b49b07767c22 · inbound

What to Keep, What to Forget: A Rate--Distortion View of Memory Compaction in LLMs and Agents cites this paper.

What to Keep, What to Forget: A Rate--Distortion View of Memory Compaction in LLMs and Agents LOOK-M: Look-Once Optimization in KV Cache for Efficient Multimodal Long-Context Inference

Reference 118

Resolution
verified exact
local_arxiv, observed 2026-07-10T01:36:44.110744Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-07-10T01:26:59.421158Z digest=sha256:ab310b6d85669d6a767c3d9041e356374f93ab15916e1b9e530bbdd21614717a

Observation f4c52329-4ebd-4876-957d-ec47b3ddc319 · inbound

MM-ShiftKV: Decode-Aware Prefill-Stage KV Selection for Multimodal Large Language Models cites this paper.

MM-ShiftKV: Decode-Aware Prefill-Stage KV Selection for Multimodal Large Language Models LOOK-M: Look-Once Optimization in KV Cache for Efficient Multimodal Long-Context Inference

Reference 79

Resolution
unresolved
no resolver link, observed 2026-08-02T11:58:29.746228Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T11:58:29.746228Z digest=sha256:a3ec567518c82fdf1707a5216e2168d9776ef5e8eb125070952290ac4ddadefa

Observation 88d04430-1dc0-4359-9dc8-82ebca331afd · inbound

PhyCheck: Fine-Grained Evidence-Grounded Dataset for Physical Law Understanding in Video-LLMs cites this paper.

PhyCheck: Fine-Grained Evidence-Grounded Dataset for Physical Law Understanding in Video-LLMs LOOK-M: Look-Once Optimization in KV Cache for Efficient Multimodal Long-Context Inference

Reference 98

Resolution
unresolved
no resolver link, observed 2026-08-04T13:43:57.570458Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T13:43:57.570458Z digest=sha256:c43df73c8f001cf067cbfda5503b4210bf00446d4fdadee7ae5d73e1916a2e52

Observation dda273cf-5b28-49da-b79b-5437f6da7fce · inbound

PhyCheck: Fine-Grained Evidence-Grounded Dataset for Physical Law Understanding in Video-LLMs cites this paper.

PhyCheck: Fine-Grained Evidence-Grounded Dataset for Physical Law Understanding in Video-LLMs LOOK-M: Look-Once Optimization in KV Cache for Efficient Multimodal Long-Context Inference

Reference 98

Resolution
unresolved
no resolver link, observed 2026-08-07T00:11:49.173076Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:11:49.173076Z digest=sha256:7a822916e0a23ed23eb1cdf10658693c62a297b00b30a6468b43de3376897fd8