Pith. sign in

Paper Citation Record · LEDGER

MME-Survey: A Comprehensive Survey on Evaluation of Multimodal LLMs

As of 11 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 33 inbound Pith citation observations for arXiv:2411.15296.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2411.15296 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 33 of 33 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00

measured 33 of 33 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-10T22:08:08.997103Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

4
arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 58b0312e-73ff-42f8-b310-a2d89b9933b8 · inbound

Visual Large Language Models for Generalized and Specialized Applications cites this paper.

Visual Large Language Models for Generalized and Specialized Applications MME-Survey: A Comprehensive Survey on Evaluation of Multimodal LLMs

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-10T22:08:08.997103Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T22:08:08.997103Z digest=sha256:74ccefcbd57c7df3735ec98ddda62fbe976c840375a8d7d7053c537fc5d6e15b

Observation d69f840a-ba74-4bf5-b900-dd7928b931eb · inbound

Position: Multimodal Large Language Models Can Significantly Advance Scientific Reasoning cites this paper.

Position: Multimodal Large Language Models Can Significantly Advance Scientific Reasoning MME-Survey: A Comprehensive Survey on Evaluation of Multimodal LLMs

Reference 46

Resolution
verified exact
arxiv_id, observed 2026-05-23T04:32:32.989284Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-23T04:30:38.804702Z digest=sha256:7b6ee2df950253c892509c72f45c3c73751b04326bf75fa37888a84b46a571a8

Observation e71ccea0-7fe7-42a3-88eb-843910d7a257 · inbound

MM-RLHF: The Next Step Forward in Multimodal LLM Alignment cites this paper.

MM-RLHF: The Next Step Forward in Multimodal LLM Alignment MME-Survey: A Comprehensive Survey on Evaluation of Multimodal LLMs

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-07T18:23:49.849204Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T18:23:49.849204Z digest=sha256:467fa00a0e986284edb339fec0868564b1bf5b59be96b90247f9ff1ca8500ae1

Observation 91af4587-f0ff-4525-8ae0-dfbce396dc6f · inbound

LENS: Multi-level Evaluation of Multimodal Reasoning with Large Language Models cites this paper.

LENS: Multi-level Evaluation of Multimodal Reasoning with Large Language Models MME-Survey: A Comprehensive Survey on Evaluation of Multimodal LLMs

Reference 9

Resolution
verified exact
arxiv_id, observed 2026-05-22T13:51:37.685457Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-22T13:47:51.436258Z digest=sha256:0fb06b7404ae6930ef81cba497e2721df9fdb4d387ae83a6022fc2ed6b053a43

Observation 5d73b1f5-bc43-45ea-8adc-38778701ecfd · inbound

STAR-R1: Spatial TrAnsformation Reasoning by Reinforcing Multimodal LLMs cites this paper.

STAR-R1: Spatial TrAnsformation Reasoning by Reinforcing Multimodal LLMs MME-Survey: A Comprehensive Survey on Evaluation of Multimodal LLMs

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-07T15:15:41.573252Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:15:41.573252Z digest=sha256:86b30d8681b434d61c0a82d06e99bb56f5915ff1bf31050bf444849322b25c7c

Observation f31ac3a0-02af-4b1c-aa28-ff6583550664 · inbound

FinMME: Benchmark Dataset for Financial Multi-Modal Reasoning Evaluation cites this paper.

FinMME: Benchmark Dataset for Financial Multi-Modal Reasoning Evaluation MME-Survey: A Comprehensive Survey on Evaluation of Multimodal LLMs

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-07T12:35:31.362263Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:35:31.362263Z digest=sha256:7b2534b1884ffc2da86d52f6d91ce1e4f2282431c64bd286cabad26bbe4ba6af

Observation 15910943-7fd6-4bd0-975b-f2f21ac59ccd · inbound

Abstractive Visual Understanding of Multi-modal Structured Knowledge: A New Perspective for MLLM Evaluation cites this paper.

Abstractive Visual Understanding of Multi-modal Structured Knowledge: A New Perspective for MLLM Evaluation MME-Survey: A Comprehensive Survey on Evaluation of Multimodal LLMs

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-07T11:51:15.288212Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:51:15.288212Z digest=sha256:e26bcdfcaac80dbf321bbe8a07a135dd00d6f01a762e32e548e31882cd52c1ba

Observation 8a236e54-b53d-4350-96ea-4fa5bbd5bdf2 · inbound

BYO-Eval: Build Your Own Dataset for Fine-Grained Visual Assessment of Multimodal Language Models cites this paper.

BYO-Eval: Build Your Own Dataset for Fine-Grained Visual Assessment of Multimodal Language Models MME-Survey: A Comprehensive Survey on Evaluation of Multimodal LLMs

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-07T10:35:40.742944Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T10:35:40.742944Z digest=sha256:4d1164a98a9fa6e7066249211a9fc5876d87c85cd43873c00a2bbdfbfbba73ab

Observation faa2101a-7e86-4794-ae79-813cfdec8a38 · inbound

Mitigating Object Hallucination via Robust Local Perception Search cites this paper.

Mitigating Object Hallucination via Robust Local Perception Search MME-Survey: A Comprehensive Survey on Evaluation of Multimodal LLMs

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-07T05:53:58.992333Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:53:58.992333Z digest=sha256:c915fb584b6bf766ded2bb244ebb0e1f1d1e32cbaa3c43cabe0c79339da329f2

Observation ceaae5c7-6498-4941-8dce-05f664adf7a8 · inbound

Hallucination at a Glance: Controlled Visual Edits and Fine-Grained Multimodal Learning cites this paper.

Hallucination at a Glance: Controlled Visual Edits and Fine-Grained Multimodal Learning MME-Survey: A Comprehensive Survey on Evaluation of Multimodal LLMs

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-07T05:43:40.351303Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:43:40.351303Z digest=sha256:66f37bd6c87677b3effd682940fa0b24168e370f2cd24494adcd98431d457502

Observation 1198782c-3978-4d7f-bd57-b5177a52c381 · inbound

Vision Generalist Model: A Survey cites this paper.

Vision Generalist Model: A Survey MME-Survey: A Comprehensive Survey on Evaluation of Multimodal LLMs

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-07T04:43:42.097080Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:43:42.097080Z digest=sha256:e1917a33d6f717facb87e2be6b8e3f8eb657902d604ff88c36d54657a5521cd0

Observation 3a649fa7-a297-4640-82a1-1d40ae8fa644 · inbound

VFaith: Do Large Multimodal Models Really Reason on Seen Images Rather than Previous Memories? cites this paper.

VFaith: Do Large Multimodal Models Really Reason on Seen Images Rather than Previous Memories? MME-Survey: A Comprehensive Survey on Evaluation of Multimodal LLMs

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-07T04:07:32.177092Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:07:32.177092Z digest=sha256:282aa732fa124c1dd4410322863580b9c09c3dc79fb9d8bd7deae89a6e0677c9

Observation b305033a-7bc9-47a0-86da-c716cee83a57 · inbound

COREVQA: A Crowd Observation and Reasoning Entailment Visual Question Answering Benchmark cites this paper.

COREVQA: A Crowd Observation and Reasoning Entailment Visual Question Answering Benchmark MME-Survey: A Comprehensive Survey on Evaluation of Multimodal LLMs

Reference 2025

Resolution
unresolved
no resolver link, observed 2026-08-06T16:41:12.474596Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:41:12.474596Z digest=sha256:68ce1d81e7d0c663046f9928fa3079f14b04b97ab6a4fd3b835e5c8a972a4a39

Observation 60c6ffac-10c6-440e-ac75-fe2fbdae59e4 · inbound

Extracting Visual Facts from Intermediate Layers for Mitigating Hallucinations in Multimodal Large Language Models cites this paper.

Extracting Visual Facts from Intermediate Layers for Mitigating Hallucinations in Multimodal Large Language Models MME-Survey: A Comprehensive Survey on Evaluation of Multimodal LLMs

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-06T15:32:34.567976Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:32:34.567976Z digest=sha256:fbc99e0c48d3362c2ab3514eaf15370ff171777e9a84dfa35e886f7ca0197cf3

Observation 25c57aa8-a239-4e56-947f-112ad31e4aa3 · inbound

LLaSO: A Foundational Framework for Reproducible Research in Large Language and Speech Model cites this paper.

LLaSO: A Foundational Framework for Reproducible Research in Large Language and Speech Model MME-Survey: A Comprehensive Survey on Evaluation of Multimodal LLMs

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-05T17:56:51.910988Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T17:56:51.910988Z digest=sha256:b8146cd6f21b9ed0f991328e8ac07344857bd314760f2c3eccb92edcb3eaf380

Observation aa0ca7aa-a81d-43a5-ba0b-89fa84c15f9e · inbound

Robix: A Unified Model for Robot Interaction, Reasoning and Planning cites this paper.

Robix: A Unified Model for Robot Interaction, Reasoning and Planning MME-Survey: A Comprehensive Survey on Evaluation of Multimodal LLMs

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-05T12:59:03.490856Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T12:59:03.490856Z digest=sha256:6975189ce7749d641dba165da16e3e8bfb12f19206515ca92ecef3ddedc9d02b

Observation 591b2eb1-bb4e-48c7-8835-189caadd707a · inbound

Structured and Abstractive Reasoning on Multi-modal Relational Knowledge Images cites this paper.

Structured and Abstractive Reasoning on Multi-modal Relational Knowledge Images MME-Survey: A Comprehensive Survey on Evaluation of Multimodal LLMs

Reference 4

Resolution
verified exact
arxiv_id, observed 2026-05-18T04:35:52.579252Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-18T04:32:54.811976Z digest=sha256:03dd5801cca23e6c6babcf0e61b02ab3a1142251fc0c96035c9a45bdf8cc71c9

Observation 44778109-6e37-4486-8e59-afb19081582d · inbound

DSBench: A Comprehensive Benchmark for Evaluating External and In-Cabin Risks cites this paper.

DSBench: A Comprehensive Benchmark for Evaluating External and In-Cabin Risks MME-Survey: A Comprehensive Survey on Evaluation of Multimodal LLMs

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-03T21:37:07.293897Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T21:37:07.293897Z digest=sha256:b757d0b185e15774979a9d09201adc56e7cb4c9af61e0e9c9a03c6cb147f6270

Observation 61c6187a-940a-49a1-a657-4ee3f4768b5d · inbound

VisPhyWorld: Probing Physical Reasoning via Code-Driven Video Reconstruction cites this paper.

VisPhyWorld: Probing Physical Reasoning via Code-Driven Video Reconstruction MME-Survey: A Comprehensive Survey on Evaluation of Multimodal LLMs

Reference 12

Resolution
verified exact
arxiv_id, observed 2026-05-22T11:11:27.456373Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-22T11:06:38.102348Z digest=sha256:ac095980e4b3f4b09c17eca12fd936d7d52bc05ac4d6922344764ddd93378c68

Observation 6389e998-af92-4262-a4fe-d0ebddf7af20 · inbound

Multi-Task Reinforcement Learning for Enhanced Multimodal LLM-as-a-Judge cites this paper.

Multi-Task Reinforcement Learning for Enhanced Multimodal LLM-as-a-Judge MME-Survey: A Comprehensive Survey on Evaluation of Multimodal LLMs

Reference 2

Resolution
verified exact
arxiv_id, observed 2026-05-15T12:35:35.313567Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-15T12:34:53.596478Z digest=sha256:f722d369c02a564240ff3520a059ddeeb196a4945568ded4bf1f90f51f9f830a

Observation 166cd3b6-082f-4e88-a2df-0d5a66694532 · inbound

Explicit Logic Channel for Validation and Enhancement of MLLMs on Zero-Shot Tasks cites this paper.

Explicit Logic Channel for Validation and Enhancement of MLLMs on Zero-Shot Tasks MME-Survey: A Comprehensive Survey on Evaluation of Multimodal LLMs

Reference 20

Resolution
verified exact
arxiv_id, observed 2026-05-21T11:10:01.996695Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-21T11:09:51.816554Z digest=sha256:21258872f2a520c19713aef065cedbe0c8a6a63a90d65316adffe26a92b800a7

Observation d053b36e-308b-474a-a990-95989a370908 · inbound

Decoding by Perturbation: Mitigating MLLM Hallucinations via Dynamic Textual Perturbation cites this paper.

Decoding by Perturbation: Mitigating MLLM Hallucinations via Dynamic Textual Perturbation MME-Survey: A Comprehensive Survey on Evaluation of Multimodal LLMs

Reference 14

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T10:51:03.774184Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-10T15:17:39.344348Z digest=sha256:169c019add8e4ef4114b141dd84cc8b7c52123a0c34187928addd5255bd07457

Observation 808d609a-0137-4df7-8089-9000bd0006b0 · inbound

Structural Ranking of the Cognitive Plausibility of Computational Models of Analogy and Metaphors with the Minimal Cognitive Grid cites this paper.

Structural Ranking of the Cognitive Plausibility of Computational Models of Analogy and Metaphors with the Minimal Cognitive Grid MME-Survey: A Comprehensive Survey on Evaluation of Multimodal LLMs

Reference 196

Resolution
verified exact
arxiv_id, observed 2026-05-11T16:56:06.127142Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-09T14:33:11.033906Z digest=sha256:f6ea994c6a058f8fec3842ff508b05ff55c6caef2bd96deb098fc1df6eced239

Observation 4f48086b-2a9e-480d-8aae-3b3209320359 · inbound

CC-OCR V2: Benchmarking Large Multimodal Models for Literacy in Real-world Document Processing cites this paper.

CC-OCR V2: Benchmarking Large Multimodal Models for Literacy in Real-world Document Processing MME-Survey: A Comprehensive Survey on Evaluation of Multimodal LLMs

Reference 65

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T23:46:49.148170Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-07T16:18:35.484800Z digest=sha256:b672e7ec3cb4e6d8c90de380c99624b73f28ad98e516117a56dfea191b9fc789

Observation 87b8f2e0-f26f-423e-afcc-d44853e0151a · inbound

SafeSteer: A Decoding-level Defense Mechanism for Multimodal Large Language Models cites this paper.

SafeSteer: A Decoding-level Defense Mechanism for Multimodal Large Language Models MME-Survey: A Comprehensive Survey on Evaluation of Multimodal LLMs

Reference 46

Resolution
metadata mismatch
arxiv_id, observed 2026-05-13T06:57:27.658510Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-13T06:56:10.053418Z digest=sha256:bb400b23b8542641c60a89243c0a4f089029bc9bf754460e8dcf84ad4ab47250

Observation c3cb4d90-09dd-4f00-87ec-90e7008fde86 · inbound

See What I Mean: Aligning Vision and Language Representations for Video Fine-grained Object Understanding cites this paper.

See What I Mean: Aligning Vision and Language Representations for Video Fine-grained Object Understanding MME-Survey: A Comprehensive Survey on Evaluation of Multimodal LLMs

Reference 21

Resolution
verified exact
arxiv_id, observed 2026-05-20T12:13:16.285669Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-20T12:10:54.874012Z digest=sha256:265cb3cdcb148547bc46af18dc7cccd07a65c8ffb3c6a884b962ad73b80a16a1

Observation f47b0468-b02e-4423-90c4-a8e7b2ee9ee3 · inbound

WinTok: A Win-Win Hybrid Tokenizer via Decomposing Visual Understanding and Generation with Transferable Tokens cites this paper.

WinTok: A Win-Win Hybrid Tokenizer via Decomposing Visual Understanding and Generation with Transferable Tokens MME-Survey: A Comprehensive Survey on Evaluation of Multimodal LLMs

Reference 31

Resolution
metadata mismatch
arxiv_id, observed 2026-05-20T12:08:15.747254Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-20T12:04:19.761430Z digest=sha256:bbc2715c5adabd35425d943f2943ab918ce4dc54aaffcafc7c8bd4453b031ea4

Observation 63ede929-1537-4a70-bc1f-a535b0984e76 · inbound

WikiVQABench: A Knowledge-Grounded Visual Question Answering Benchmark from Wikipedia and Wikidata cites this paper.

WikiVQABench: A Knowledge-Grounded Visual Question Answering Benchmark from Wikipedia and Wikidata MME-Survey: A Comprehensive Survey on Evaluation of Multimodal LLMs

Reference 115

Resolution
metadata mismatch
arxiv_id, observed 2026-05-21T04:43:58.708522Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-21T04:41:06.779529Z digest=sha256:bb503b3b93ece90afe93c4af373d64badf26dac99260b717c708389dba905bec

Observation 1e5eda72-809b-41a2-97cb-c4920fb07921 · inbound

LLM-Based Examination of Eligibility Criteria from Securities Prospectuses at the German Central Bank cites this paper.

LLM-Based Examination of Eligibility Criteria from Securities Prospectuses at the German Central Bank MME-Survey: A Comprehensive Survey on Evaluation of Multimodal LLMs

Reference 24

Resolution
metadata mismatch
arxiv_id, observed 2026-06-26T03:58:57.116684Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-06-26T03:56:28.271760Z digest=sha256:7540eb924d94028fa183a1e31656bafd47f2fad0bacfad31133a0bbf1c7451df

Observation 52278575-45e4-495c-8025-2304dd4f967b · inbound

Causal Connections: Leveraging Multilingual Fine-Tuning for Financial QA@FinCausal 2026 cites this paper.

Causal Connections: Leveraging Multilingual Fine-Tuning for Financial QA@FinCausal 2026 MME-Survey: A Comprehensive Survey on Evaluation of Multimodal LLMs

Reference 54

Resolution
metadata mismatch
arxiv_id, observed 2026-06-29T02:23:00.931773Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-06-29T02:17:26.872854Z digest=sha256:5272bae450de53142107b656d76ecec922dcd9c6b55783b977ffda47832b12ad

Observation d3532870-03af-48ba-8a64-349c3088724e · inbound

Generalize LMMs to Versatile Visual Modalities via Fabricated Modality Synthesis cites this paper.

Generalize LMMs to Versatile Visual Modalities via Fabricated Modality Synthesis MME-Survey: A Comprehensive Survey on Evaluation of Multimodal LLMs

Reference 10

Resolution
unresolved
no resolver link, observed 2026-07-14T12:45:30.225562Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T12:45:30.225562Z digest=sha256:1d1245f58a761e5c2ee2bc2e9e85cb7abcd2e8ac8d5c5c551134e3425b3c272a

Observation 5cbdc1b8-f069-4838-a1cb-7d9a07af4024 · inbound

Contextualized Evaluation of Vision Language Models through Dynamic, Multi-turn Interactions cites this paper.

Contextualized Evaluation of Vision Language Models through Dynamic, Multi-turn Interactions MME-Survey: A Comprehensive Survey on Evaluation of Multimodal LLMs

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-02T01:58:01.941030Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T01:58:01.941030Z digest=sha256:79ec01752c4fad32e239d85c005914b2c85cf66a66ae4bfe2ae40e2df9f8d084

Observation 1a816af5-5c56-4df6-86ff-b56fda2106c0 · inbound

SABRE: Scalable and Automated Benchmarking of VLMs under Stress cites this paper.

SABRE: Scalable and Automated Benchmarking of VLMs under Stress MME-Survey: A Comprehensive Survey on Evaluation of Multimodal LLMs

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-10T04:45:07.223493Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T04:45:07.223493Z digest=sha256:21fc25b5cecdcded7597ec67110fd2110c0aa48ac109915112af3709cd638738