Pith. sign in

Paper Citation Record · LEDGER

Mementos: A Comprehensive Benchmark for Multimodal Large Language Model Reasoning over Image Sequences

As of 4 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 7 inbound Pith citation observations for arXiv:2401.10529.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2401.10529 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 7 of 7 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-04T06:34:03.388597+00:00

measured 7 of 7 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-03T23:35:18.926817Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-20T19:13:40.522467Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 90f2d030-8bf9-4556-ab50-675d6e5521a9 · inbound

MuirBench: A Comprehensive Benchmark for Robust Multi-image Understanding cites this paper.

MuirBench: A Comprehensive Benchmark for Robust Multi-image Understanding Mementos: A Comprehensive Benchmark for Multimodal Large Language Model Reasoning over Image Sequences

Reference 65

Resolution
verified exact
arxiv_id, observed 2026-05-17T01:09:30.446521Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-17T01:09:30.360275Z digest=sha256:65d751dd8a37ca83dec891b62338c7590274c32541af16c6c471ac24813d7ae0

Observation 04cca4ad-d10f-4e47-966c-7749221d48db · inbound

mPLUG-Owl3: Towards Long Image-Sequence Understanding in Multi-Modal Large Language Models cites this paper.

mPLUG-Owl3: Towards Long Image-Sequence Understanding in Multi-Modal Large Language Models Mementos: A Comprehensive Benchmark for Multimodal Large Language Model Reasoning over Image Sequences

Reference 118

Resolution
metadata mismatch
arxiv_id, observed 2026-05-20T06:20:36.548843Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-20T06:20:36.235304Z digest=sha256:6eefb71d57c5a020273a4cb75a95bac793494df8844ccd2c1cd0753de77535e4

Observation f2f9a9dc-f0df-4599-93ea-7c7e563967cc · inbound

Expanding Performance Boundaries of Open-Source Multimodal Models with Model, Data, and Test-Time Scaling cites this paper.

Expanding Performance Boundaries of Open-Source Multimodal Models with Model, Data, and Test-Time Scaling Mementos: A Comprehensive Benchmark for Multimodal Large Language Model Reasoning over Image Sequences

Reference 255

Resolution
verified exact
arxiv_id, observed 2026-05-10T13:23:58.136717Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-10T13:23:57.588851Z digest=sha256:377802578d70464f800135c342bcb2dd6ec6269e16abfc1bea614d6902c60a0b

Observation 8b18b037-8b02-4f4e-86ca-244c91950496 · inbound

Towards Mitigating Hallucinations in Large Vision-Language Models by Refining Textual Embeddings cites this paper.

Towards Mitigating Hallucinations in Large Vision-Language Models by Refining Textual Embeddings Mementos: A Comprehensive Benchmark for Multimodal Large Language Model Reasoning over Image Sequences

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-03T23:35:18.926817Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T23:35:18.926817Z digest=sha256:7b2bbf54c649c8c97706c1be011ec05f0d0629f47177175c525eb47c85270f15

Observation a0c0d88c-c8f5-4176-9148-54c8e93ec1ad · inbound

Spatio-Temporal Grounding of Large Language Models from Perception Streams cites this paper.

Spatio-Temporal Grounding of Large Language Models from Perception Streams Mementos: A Comprehensive Benchmark for Multimodal Large Language Model Reasoning over Image Sequences

Reference 25

Resolution
verified exact
arxiv_id, observed 2026-05-11T07:26:02.674180Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-10T17:10:45.837684Z digest=sha256:e127eff40513ba6d81b5ec73468464db3c391d56696c514fe5149e166fea7412

Observation 2583d57f-3e94-417a-bda6-ee3a603071d1 · inbound

SMMBench: A Benchmark for Source-Distributed Multimodal Agent Memory cites this paper.

SMMBench: A Benchmark for Source-Distributed Multimodal Agent Memory Mementos: A Comprehensive Benchmark for Multimodal Large Language Model Reasoning over Image Sequences

Reference 12

Resolution
verified exact
arxiv_id, observed 2026-05-20T19:13:40.524351Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-20T19:11:15.761831Z digest=sha256:c14a133b785ee08479c4406d3cf21dd20d676660e9872d120a8adc64609395b1

Observation f5853c1d-4920-472a-b56f-c7429762a3f2 · inbound

Beyond Retrieval: Analytic Memory for Multimodal Agents cites this paper.

Beyond Retrieval: Analytic Memory for Multimodal Agents Mementos: A Comprehensive Benchmark for Multimodal Large Language Model Reasoning over Image Sequences

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-03T07:06:05.783875Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-03T07:06:05.783875Z digest=sha256:a10282bc9f6284265da29b40466c5500beecf1c03c14cc4df49d87dd8639fa8c