Pith. sign in

Paper Citation Record · LEDGER

MRAG-Bench: Vision-Centric Evaluation for Retrieval-Augmented Multimodal Models

As of 22 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 12 inbound Pith citation observations for arXiv:2410.08182.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2410.08182 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 12 of 12 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-21T06:32:19.484+00:00

measured 12 of 12 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-16T11:38:48.463860Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-06-30T21:05:04.211824Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation d8580dc8-f4c2-4272-940c-64a4fbe28cd6 · inbound

The Great Nugget Recall: Automating Fact Extraction and RAG Evaluation with Large Language Models cites this paper.

The Great Nugget Recall: Automating Fact Extraction and RAG Evaluation with Large Language Models MRAG-Bench: Vision-Centric Evaluation for Retrieval-Augmented Multimodal Models

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-16T11:38:48.463860Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:38:48.463860Z digest=sha256:f685489a89da746fb8f55d408212253a346de0a40260ce97e920fdc303376b5f

Observation 53986ec4-e589-4e80-a516-5f5bbe5bfa7a · inbound

Towards General Continuous Memory for Vision-Language Models cites this paper.

Towards General Continuous Memory for Vision-Language Models MRAG-Bench: Vision-Centric Evaluation for Retrieval-Augmented Multimodal Models

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-07T14:46:46.436410Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:46:46.436410Z digest=sha256:4941e21c791285b767a43e191a34de6ef0a4d2838011e788faf3e0b877d30914

Observation 833d612a-f0b1-4e38-9573-6800cfc0396f · inbound

VisRet: Visualization Improves Knowledge-Intensive Text-to-Image Retrieval cites this paper.

VisRet: Visualization Improves Knowledge-Intensive Text-to-Image Retrieval MRAG-Bench: Vision-Centric Evaluation for Retrieval-Augmented Multimodal Models

Reference 4

Resolution
metadata mismatch
arxiv_id, observed 2026-05-19T12:52:18.113553Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-19T12:48:41.768247Z digest=sha256:5ed5232bfc4a6e8e957cb0c5a41053f73eae9b12cfd12062f6ff896955cd93d5

Observation 919e3d3b-3219-4415-bf1e-1c94fb452c1f · inbound

R3G: A Reasoning-Retrieval-Reranking Framework for Vision-Centric Answer Generation cites this paper.

R3G: A Reasoning-Retrieval-Reranking Framework for Vision-Centric Answer Generation MRAG-Bench: Vision-Centric Evaluation for Retrieval-Augmented Multimodal Models

Reference 18

Resolution
metadata mismatch
arxiv_id, observed 2026-05-16T10:47:45.527711Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-16T10:46:17.411843Z digest=sha256:484d55689a1421cf6858147b193a22ae142796050cf9d16155ff1a766cee9ca8

Observation e03881ca-6b7a-4a28-8806-24b7e7ef23c7 · inbound

R3G: A Reasoning-Retrieval-Reranking Framework for Vision-Centric Answer Generation cites this paper.

R3G: A Reasoning-Retrieval-Reranking Framework for Vision-Centric Answer Generation MRAG-Bench: Vision-Centric Evaluation for Retrieval-Augmented Multimodal Models

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-03T08:13:52.855031Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T08:13:52.855031Z digest=sha256:7c888fc677c92629b63ee6e818b783f1aba2e62d658e1f8202b9e957bee8c7be

Observation 4e5ef75b-7586-4e2e-ba03-6839fe95f17d · inbound

CoGR-MoE: Concept-Guided Expert Routing with Consistent Selection and Flexible Reasoning for Visual Question Answering cites this paper.

CoGR-MoE: Concept-Guided Expert Routing with Consistent Selection and Flexible Reasoning for Visual Question Answering MRAG-Bench: Vision-Centric Evaluation for Retrieval-Augmented Multimodal Models

Reference 11

Resolution
verified exact
arxiv_id, observed 2026-05-10T07:11:53.233363Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-05-10T07:09:48.239662Z digest=sha256:25b151a73c00b0f0055251886f63ef89b89d97c3dce26d7fcffc18962737042a

Observation 064e9d8b-61f3-4954-b41b-1fc14b586699 · inbound

Utility-Oriented Visual Evidence Selection for Multimodal Retrieval-Augmented Generation cites this paper.

Utility-Oriented Visual Evidence Selection for Multimodal Retrieval-Augmented Generation MRAG-Bench: Vision-Centric Evaluation for Retrieval-Augmented Multimodal Models

Reference 11

Resolution
verified exact
arxiv_id, observed 2026-05-14T19:09:22.883718Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-05-14T19:09:18.975682Z digest=sha256:f2b81475a0614aea09e4928e8ebe4df41c9a01d55f0a9b903bbebc2f2d0b53a2

Observation f54609e6-3d1a-4115-99e9-34d0e9999825 · inbound

MemLens: Benchmarking Multimodal Long-Term Memory in Large Vision-Language Models cites this paper.

MemLens: Benchmarking Multimodal Long-Term Memory in Large Vision-Language Models MRAG-Bench: Vision-Centric Evaluation for Retrieval-Augmented Multimodal Models

Reference 20

Resolution
verified exact
arxiv_id, observed 2026-06-30T21:05:04.213305Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-06-30T21:00:25.664841Z digest=sha256:c7866c5a1c360ed1bd21149f7ec05f95c3653ade665efe2b847549dfa1f28fb2

Observation 0d7687f8-b817-4f41-98ab-b7334465d62c · inbound

SVFSearch: A Multimodal Knowledge-Intensive Benchmark for Short-Video Frame Search in the Gaming Vertical Domain cites this paper.

SVFSearch: A Multimodal Knowledge-Intensive Benchmark for Short-Video Frame Search in the Gaming Vertical Domain MRAG-Bench: Vision-Centric Evaluation for Retrieval-Augmented Multimodal Models

Reference 53

Resolution
verified exact
arxiv_id, observed 2026-05-20T10:13:11.985883Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-05-20T10:08:56.397296Z digest=sha256:4871810be6698d567b11197c5f1ad94bd6573eaa98aa98a39b2bb1f21f36287f

Observation eee3052f-916e-4e8d-80ed-5428559a68b2 · inbound

SVFSearch: A Multimodal Knowledge-Intensive Benchmark for Short-Video Frame Search in the Gaming Vertical Domain cites this paper.

SVFSearch: A Multimodal Knowledge-Intensive Benchmark for Short-Video Frame Search in the Gaming Vertical Domain MRAG-Bench: Vision-Centric Evaluation for Retrieval-Augmented Multimodal Models

Reference 53

Resolution
verified exact
arxiv_id, observed 2026-05-21T08:49:53.437595Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-05-21T08:49:32.503461Z digest=sha256:64124117a58a0b5275e63408641831b32e0530f64a60ada37634fdaaa8c83468

Observation 5dea582f-aeff-46db-b45c-835b9157dbdb · inbound

Reason Before You Retrieve: Agentic Planning for Multi-modal RAG cites this paper.

Reason Before You Retrieve: Agentic Planning for Multi-modal RAG MRAG-Bench: Vision-Centric Evaluation for Retrieval-Augmented Multimodal Models

Reference 136

Resolution
unresolved
no resolver link, observed 2026-08-02T10:20:56.184193Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T10:20:56.184193Z digest=sha256:aca4a86b008448ee919106572b85f00129256fceb7b9ac0090a18c57ac3dd45a

Observation 8ea05033-810f-4109-8777-0960d41aebdd · inbound

VLD-RAG: Agentic Vision-Language Retrieval-Augmented Generation for Long, Visually-Rich Multi-Page Documents cites this paper.

VLD-RAG: Agentic Vision-Language Retrieval-Augmented Generation for Long, Visually-Rich Multi-Page Documents MRAG-Bench: Vision-Centric Evaluation for Retrieval-Augmented Multimodal Models

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-02T13:26:29.810681Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T13:26:29.810681Z digest=sha256:e43fda69fa81113bd4f867c0b13c3a9ab31583b96dff3aa1d27637a2f68f8c58