Pith. sign in

Paper Citation Record · LEDGER

MRAG-Bench: Vision-Centric Evaluation for Retrieval-Augmented Multimodal Models

As of 8 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 11 inbound Pith citation observations for arXiv:2410.08182.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2410.08182 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 11 of 11 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 11 of 11 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T14:46:46.436410Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-06-30T21:05:04.211824Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 53986ec4-e589-4e80-a516-5f5bbe5bfa7a · inbound

Towards General Continuous Memory for Vision-Language Models cites this paper.

Towards General Continuous Memory for Vision-Language Models MRAG-Bench: Vision-Centric Evaluation for Retrieval-Augmented Multimodal Models

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-07T14:46:46.436410Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:46:46.436410Z digest=sha256:414ce6e2241a4184abebe3368da0387492f1e455337597e58c8fa6ba1668a383

Observation 833d612a-f0b1-4e38-9573-6800cfc0396f · inbound

VisRet: Visualization Improves Knowledge-Intensive Text-to-Image Retrieval cites this paper.

VisRet: Visualization Improves Knowledge-Intensive Text-to-Image Retrieval MRAG-Bench: Vision-Centric Evaluation for Retrieval-Augmented Multimodal Models

Reference 4

Resolution
metadata mismatch
arxiv_id, observed 2026-05-19T12:52:18.113553Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-19T12:48:41.768247Z digest=sha256:0c97d5348147dbfda82dea767ef4435bc5b94ace27631f4ed53b9707b077fe64

Observation 919e3d3b-3219-4415-bf1e-1c94fb452c1f · inbound

R3G: A Reasoning-Retrieval-Reranking Framework for Vision-Centric Answer Generation cites this paper.

R3G: A Reasoning-Retrieval-Reranking Framework for Vision-Centric Answer Generation MRAG-Bench: Vision-Centric Evaluation for Retrieval-Augmented Multimodal Models

Reference 18

Resolution
metadata mismatch
arxiv_id, observed 2026-05-16T10:47:45.527711Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-16T10:46:17.411843Z digest=sha256:4b8015a6d1e410e1bf5c49d5ac5bd57d7cf9439048dcbd2550151a5922bcf1f2

Observation e03881ca-6b7a-4a28-8806-24b7e7ef23c7 · inbound

R3G: A Reasoning-Retrieval-Reranking Framework for Vision-Centric Answer Generation cites this paper.

R3G: A Reasoning-Retrieval-Reranking Framework for Vision-Centric Answer Generation MRAG-Bench: Vision-Centric Evaluation for Retrieval-Augmented Multimodal Models

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-03T08:13:52.855031Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T08:13:52.855031Z digest=sha256:474850fe8f689139ed8fbbe720e6a81fed193ac3b2d72126ffade380c198a5cd

Observation 4e5ef75b-7586-4e2e-ba03-6839fe95f17d · inbound

CoGR-MoE: Concept-Guided Expert Routing with Consistent Selection and Flexible Reasoning for Visual Question Answering cites this paper.

CoGR-MoE: Concept-Guided Expert Routing with Consistent Selection and Flexible Reasoning for Visual Question Answering MRAG-Bench: Vision-Centric Evaluation for Retrieval-Augmented Multimodal Models

Reference 11

Resolution
verified exact
arxiv_id, observed 2026-05-10T07:11:53.233363Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-05-10T07:09:48.239662Z digest=sha256:4e9a2e110a3c1ffa6a64c92e07884e3ae2d28a9764da4bcdf997eb8bdb6b0f54

Observation 064e9d8b-61f3-4954-b41b-1fc14b586699 · inbound

Utility-Oriented Visual Evidence Selection for Multimodal Retrieval-Augmented Generation cites this paper.

Utility-Oriented Visual Evidence Selection for Multimodal Retrieval-Augmented Generation MRAG-Bench: Vision-Centric Evaluation for Retrieval-Augmented Multimodal Models

Reference 11

Resolution
verified exact
arxiv_id, observed 2026-05-14T19:09:22.883718Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-05-14T19:09:18.975682Z digest=sha256:0acf86a3d1c8ec71ca51b7884fd89a0b68160f65903cd9db5c0a7224a1be0768

Observation f54609e6-3d1a-4115-99e9-34d0e9999825 · inbound

MemLens: Benchmarking Multimodal Long-Term Memory in Large Vision-Language Models cites this paper.

MemLens: Benchmarking Multimodal Long-Term Memory in Large Vision-Language Models MRAG-Bench: Vision-Centric Evaluation for Retrieval-Augmented Multimodal Models

Reference 20

Resolution
verified exact
arxiv_id, observed 2026-06-30T21:05:04.213305Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-30T21:00:25.664841Z digest=sha256:f3251f6acb8465f228444594814c1790ff4d85e4bc563956a3c81f2002b08e77

Observation 0d7687f8-b817-4f41-98ab-b7334465d62c · inbound

SVFSearch: A Multimodal Knowledge-Intensive Benchmark for Short-Video Frame Search in the Gaming Vertical Domain cites this paper.

SVFSearch: A Multimodal Knowledge-Intensive Benchmark for Short-Video Frame Search in the Gaming Vertical Domain MRAG-Bench: Vision-Centric Evaluation for Retrieval-Augmented Multimodal Models

Reference 53

Resolution
verified exact
arxiv_id, observed 2026-05-20T10:13:11.985883Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-05-20T10:08:56.397296Z digest=sha256:51010d4bead2910f0ce3a56c87fbdc1035308633550ad9d9c4f3217b5591258c

Observation eee3052f-916e-4e8d-80ed-5428559a68b2 · inbound

SVFSearch: A Multimodal Knowledge-Intensive Benchmark for Short-Video Frame Search in the Gaming Vertical Domain cites this paper.

SVFSearch: A Multimodal Knowledge-Intensive Benchmark for Short-Video Frame Search in the Gaming Vertical Domain MRAG-Bench: Vision-Centric Evaluation for Retrieval-Augmented Multimodal Models

Reference 53

Resolution
verified exact
arxiv_id, observed 2026-05-21T08:49:53.437595Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-05-21T08:49:32.503461Z digest=sha256:d3227cee0ee738defaaa5f8582fcb454b46ac8818489864adc3f146c2f9a54c2

Observation 5dea582f-aeff-46db-b45c-835b9157dbdb · inbound

Reason Before You Retrieve: Agentic Planning for Multi-modal RAG cites this paper.

Reason Before You Retrieve: Agentic Planning for Multi-modal RAG MRAG-Bench: Vision-Centric Evaluation for Retrieval-Augmented Multimodal Models

Reference 136

Resolution
unresolved
no resolver link, observed 2026-08-02T10:20:56.184193Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T10:20:56.184193Z digest=sha256:d755c3d3a136f53d70a5a50fd4e56c28afe8a5d4dcaba3d5c7082eb2dc9a4f05

Observation 8ea05033-810f-4109-8777-0960d41aebdd · inbound

VLD-RAG: Agentic Vision-Language Retrieval-Augmented Generation for Long, Visually-Rich Multi-Page Documents cites this paper.

VLD-RAG: Agentic Vision-Language Retrieval-Augmented Generation for Long, Visually-Rich Multi-Page Documents MRAG-Bench: Vision-Centric Evaluation for Retrieval-Augmented Multimodal Models

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-02T13:26:29.810681Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T13:26:29.810681Z digest=sha256:807d7ff56498f648070fb6e197ef523a2ed796b15833a3f04b874bf22b7597b8