Pith. sign in

Paper Citation Record · LEDGER

MileBench: Benchmarking MLLMs in Long Context

As of 8 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 18 inbound Pith citation observations for arXiv:2404.18532.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2404.18532 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 18 of 18 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 18 of 18 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T14:33:32.170655Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-04T13:19:50.957441Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 54251bc8-7ba2-4f9c-bd6f-1ac1e72d4c61 · inbound

Rethinking Causal Mask Attention for Vision-Language Inference cites this paper.

Rethinking Causal Mask Attention for Vision-Language Inference MileBench: Benchmarking MLLMs in Long Context

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-07T14:33:32.170655Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:33:32.170655Z digest=sha256:420c8e61c8221fec37b6ada312b95960adbc96a0a194b1cf9a53d93a80bc1109

Observation 0d9ba4f4-81b9-478f-84f5-54dd0f3419fd · inbound

Evaluating MLLMs with Multimodal Multi-image Reasoning Benchmark cites this paper.

Evaluating MLLMs with Multimodal Multi-image Reasoning Benchmark MileBench: Benchmarking MLLMs in Long Context

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-07T11:05:19.523326Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:05:19.523326Z digest=sha256:188590c3a0704d3aff1bde7af0410318c137380366073b34140b0bf6f78c3f9f

Observation d554ce3f-0191-4126-aca7-caf3d993cddc · inbound

CoMemo: LVLMs Need Image Context with Image Memory cites this paper.

CoMemo: LVLMs Need Image Context with Image Memory MileBench: Benchmarking MLLMs in Long Context

Reference 90

Resolution
unresolved
no resolver link, observed 2026-08-07T06:02:30.430076Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T06:02:30.430076Z digest=sha256:b7f29bae9fc19a4a7646de9f95e9548cad722e0d3a980fdaeb12ebd8dd44603d

Observation 880fad2a-01fb-4971-8831-b07aabac3f1b · inbound

MadaKV: Adaptive Modality-Perception KV Cache Eviction for Efficient Multimodal Long-Context Inference cites this paper.

MadaKV: Adaptive Modality-Perception KV Cache Eviction for Efficient Multimodal Long-Context Inference MileBench: Benchmarking MLLMs in Long Context

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-07T10:20:43.291216Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:20:43.291216Z digest=sha256:79163eeffefcfc53e41f8a0106f5184e8e552f0d589b4e714b25a0a164782854

Observation 2c6eae63-2f9c-449f-b856-6080ae010be5 · inbound

Document Haystack: A Long Context Multimodal Image/Document Understanding Vision LLM Benchmark cites this paper.

Document Haystack: A Long Context Multimodal Image/Document Understanding Vision LLM Benchmark MileBench: Benchmarking MLLMs in Long Context

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-06T16:12:02.171862Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:12:02.171862Z digest=sha256:3ea1a47f1f080ce0acf3fdbe13eec48be0a19312cf2879bd44676a3a25f32a03

Observation 201b3a3c-b695-4954-b796-3d4b35a02981 · inbound

Detecting Hope, Hate, and Emotion in Arabic Textual Speech and Multi-modal Memes Using Large Language Models cites this paper.

Detecting Hope, Hate, and Emotion in Arabic Textual Speech and Multi-modal Memes Using Large Language Models MileBench: Benchmarking MLLMs in Long Context

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-05T20:02:29.383622Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:02:29.383622Z digest=sha256:dbf511c38bb55785be25b107e2e3a0dcdb3c74e367bd7c8d4e307439b01c86f7

Observation 3af10797-c208-4084-b5cd-a7c05ee04a06 · inbound

HybridKV: Hybrid KV Cache Compression for Efficient Multimodal Large Language Model Inference cites this paper.

HybridKV: Hybrid KV Cache Compression for Efficient Multimodal Large Language Model Inference MileBench: Benchmarking MLLMs in Long Context

Reference 2

Resolution
metadata mismatch
arxiv_id, observed 2026-05-10T22:20:48.365349Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-10T19:58:45.374139Z digest=sha256:eb5445e35be47686757d63997bfd2f9fe01dc0830ef2c44cc60ccd418654044d

Observation a0cc7951-5cd0-430c-b556-f64234d04640 · inbound

COHERENCE: Benchmarking Fine-Grained Image-Text Alignment in Interleaved Multimodal Contexts cites this paper.

COHERENCE: Benchmarking Fine-Grained Image-Text Alignment in Interleaved Multimodal Contexts MileBench: Benchmarking MLLMs in Long Context

Reference 39

Resolution
verified exact
arxiv_id, observed 2026-05-12T09:41:26.768262Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-07T09:47:27.891288Z digest=sha256:1113c0d58694e6d9652a59d506b53137ea438115652a39f0cd6ae1a6fa077da6

Observation a4889efe-5619-4e05-85a1-fb03afec0c7e · inbound

COHERENCE: Benchmarking Fine-Grained Image-Text Alignment in Interleaved Multimodal Contexts cites this paper.

COHERENCE: Benchmarking Fine-Grained Image-Text Alignment in Interleaved Multimodal Contexts MileBench: Benchmarking MLLMs in Long Context

Reference 39

Resolution
verified exact
arxiv_id, observed 2026-05-14T21:17:59.250170Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-14T21:15:30.952247Z digest=sha256:3d9edf6178b81a422f80a5fe36d9e34ea477ca3a8809dc85f5675764d7a061ad

Observation 3bf906e0-6e71-48f3-98aa-5f0ec809dcd8 · inbound

WindowQuant: Mixed-Precision KV Cache Quantization based on Window-Level Similarity for VLMs Inference Optimization cites this paper.

WindowQuant: Mixed-Precision KV Cache Quantization based on Window-Level Similarity for VLMs Inference Optimization MileBench: Benchmarking MLLMs in Long Context

Reference 38

Resolution
verified exact
arxiv_id, observed 2026-05-11T16:36:07.948947Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-09T16:06:26.450483Z digest=sha256:a365e2ba3b9aca75ec5719b01b67a9d610fd943b2f6f8b9703b29878643acdce

Observation 8bca502f-b286-4aa9-87eb-9325ef1d5719 · inbound

RetentiveKV: State-Space Memory for Uncertainty-Aware Multimodal KV Cache Eviction cites this paper.

RetentiveKV: State-Space Memory for Uncertainty-Aware Multimodal KV Cache Eviction MileBench: Benchmarking MLLMs in Long Context

Reference 48

Resolution
verified exact
arxiv_id, observed 2026-05-11T10:26:02.126197Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-05-10T15:29:17.567557Z digest=sha256:86dcb531b00f09f2241050e2b632cacb2e0f5cf13abd821d21f20a83f2c28e07

Observation 117c5620-2cee-4b79-8911-f7da3513b579 · inbound

Logit-Attention Divergence: Mitigating Position Bias in Multi-Image Retrieval via Attention-Guided Calibration cites this paper.

Logit-Attention Divergence: Mitigating Position Bias in Multi-Image Retrieval via Attention-Guided Calibration MileBench: Benchmarking MLLMs in Long Context

Reference 5

Resolution
metadata mismatch
arxiv_id, observed 2026-05-13T01:27:01.748004Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-13T01:26:45.182597Z digest=sha256:75f88b7cb63859f026ae410a2d514740cec51e204cd535f2579e3477272b80b2

Observation 5a435841-c8b9-4dfc-ad51-0309e22c2e8b · inbound

MMCL-Bench: Multimodal Context Learning from Visual Rules, Procedures, and Evidence cites this paper.

MMCL-Bench: Multimodal Context Learning from Visual Rules, Procedures, and Evidence MileBench: Benchmarking MLLMs in Long Context

Reference 6

Resolution
verified exact
arxiv_id, observed 2026-05-14T20:59:27.482431Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-14T20:55:50.873271Z digest=sha256:2ad39d48a9ed7e379e76aa9f1bbb7228ba3a424d186aa8713507f512cfafcdab

Observation 739ec28f-6829-47d1-9e38-074a0d7ff60c · inbound

MemLens: Benchmarking Multimodal Long-Term Memory in Large Vision-Language Models cites this paper.

MemLens: Benchmarking Multimodal Long-Term Memory in Large Vision-Language Models MileBench: Benchmarking MLLMs in Long Context

Reference 56

Resolution
verified exact
arxiv_id, observed 2026-06-30T21:05:04.279043Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-30T21:00:25.664841Z digest=sha256:2e58234251aab1c0d1f6dd6ab6c7d025f332fb2d2c7ce1e285ebf888f7100dc1

Observation 6f75a050-2611-44a5-b3b8-623a2d38aa37 · inbound

SMMBench: A Benchmark for Source-Distributed Multimodal Agent Memory cites this paper.

SMMBench: A Benchmark for Source-Distributed Multimodal Agent Memory MileBench: Benchmarking MLLMs in Long Context

Reference 11

Resolution
verified exact
arxiv_id, observed 2026-05-20T19:13:40.544030Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-20T19:11:15.761831Z digest=sha256:0d567a8ce22571851f16473730c79181051fcf6549a05080a3bf0c4acc482b2f

Observation ad7d0583-4333-40dd-ba45-118a7ff0333c · inbound

DiCoBench: Benchmarking Multi-Image Fine-Grained Perception via Differential and Commonality Visual Cues cites this paper.

DiCoBench: Benchmarking Multi-Image Fine-Grained Perception via Differential and Commonality Visual Cues MileBench: Benchmarking MLLMs in Long Context

Reference 22

Resolution
metadata mismatch
arxiv_id, observed 2026-07-04T13:19:50.958652Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-26T05:18:01.931929Z digest=sha256:c5f2439bdd16e75c282511918321f9e85ce65f5145f129d031627c1fde0cf956

Observation a8857015-3a1b-4518-9281-64cdbda5985d · inbound

CLBench-V: Evaluating Multimodal Context Learning from Grounding to Knowledge Acquisition cites this paper.

CLBench-V: Evaluating Multimodal Context Learning from Grounding to Knowledge Acquisition MileBench: Benchmarking MLLMs in Long Context

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-01T02:55:23.417862Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T02:55:23.417862Z digest=sha256:43320f92f0e64f4dd93498bc82685139809c9483b54ec5a65dcbdb1c089e52c6

Observation 49814009-ab6d-43c4-b880-db4335dddc74 · inbound

PMMC: Prospective Multimodal Memory Compilation for Long-Term LVLM Agents cites this paper.

PMMC: Prospective Multimodal Memory Compilation for Long-Term LVLM Agents MileBench: Benchmarking MLLMs in Long Context

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-06T00:41:02.947856Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T00:41:02.947856Z digest=sha256:f2a71e08f1313930feac86568ad4b81ac193f98c01557cc1fc47aeaadfb7f199