Pith. sign in

Paper Citation Record · LEDGER

MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM

As of 22 August 2026, this Paper Citation Record lists 87 of 87 outbound references and 4 inbound Pith citation observations for arXiv:2505.24238.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.24238 v2

Coverage vector

measured 87 of 87 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T12:33:13.820971Z

measured 91 of 91 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-21T06:32:19.484+00:00

measured 4 of 4 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-05T10:39:03.028373Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Reference resolution

87 of 87 outbound references displayed

  • verified exact1
  • verified fuzzy18
  • unresolved67
  • parse uncertain0
  • malformed identifier1
  • metadata mismatch0

External citation measurements

0
arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Outbound references

Observation cd373f87-c04c-4530-bc32-e4af6da0f6fd · outbound

This paper cites Phi-3 Technical Report: A Highly Capable Language Model Locally on Your Phone.

MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM Phi-3 Technical Report: A Highly Capable Language Model Locally on Your Phone

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-07T12:33:06.080923Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:33:06.080923Z digest=sha256:4aef3fb29c813af1336ccccc95969119690dda1491b039257edb45d30df7ef24

Observation 1f5eb507-8925-4c9a-b2f3-080a49e013e3 · outbound

This paper cites Qwen2.5-VL Technical Report.

MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM Qwen2.5-VL Technical Report

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-07T12:33:06.122153Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:33:06.122153Z digest=sha256:7da9dd9cc9ceee9066a488307a44693cd71da6f3bc366b4f4e6706947a33f55d

Observation e235ea63-d9dd-4ac5-bd0b-e70324949e95 · outbound

This paper cites Mitigating Open-Vocabulary Caption Hallucinations.

MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM Mitigating Open-Vocabulary Caption Hallucinations

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-07T12:33:06.187286Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:33:06.187286Z digest=sha256:3bf8c075a14c1d3729e8f2a50297e883e8bea0f29261b2059107bb348dd96ac0

Observation 0ac1d8d1-47ac-46b8-a4fa-ba1fb0af3714 · outbound

This paper cites MM-IQ: Benchmarking Human-Like Abstraction and Reasoning in Multimodal Models.

MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM MM-IQ: Benchmarking Human-Like Abstraction and Reasoning in Multimodal Models

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-07T12:33:06.268894Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:33:06.268894Z digest=sha256:fe27c9fec7dc94df20ec817b5e726efc05e724bfcf9d18b26c0e465714a2dca0

Observation a02ce98e-e834-4a5d-8358-f84a06c5ca25 · outbound

This paper cites Internvl: Scaling up vision foundation models and aligning for generic visual-linguistic tasks.

MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM Internvl: Scaling up vision foundation models and aligning for generic visual-linguistic tasks

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-07T12:33:06.382734Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:33:06.382734Z digest=sha256:621656e1e8c5d845e72012c2a60950ffd573dd290091502edc0f32ccd8744f72

Observation 6b2a2fc9-cbcb-4800-9df7-e62cbf88b301 · outbound

This paper cites Mitigating Hallucination in Visual Language Models with Visual Supervision.

MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM Mitigating Hallucination in Visual Language Models with Visual Supervision

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-07T12:33:06.475621Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:33:06.475621Z digest=sha256:7b7cc74dfafa15f322ba8d6aadecd5ee01157dd70bdbfdc69fc193dfd4a5c90a

Observation 60e44d62-1f95-4794-bbb8-ed07a8806100 · outbound

This paper cites Puzzlevqa: Diagnosing multimodal reasoning challenges of language models with abstract visual patterns.

MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM Puzzlevqa: Diagnosing multimodal reasoning challenges of language models with abstract visual patterns

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-07T12:33:06.558504Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:33:06.558504Z digest=sha256:fee012930301bccce7b898acfd7f355d0ec51e0a23f147d3e3161c52b64fd36d

Observation e6c82916-4191-463e-9612-10cbbf6c873d · outbound

This paper cites Zero-shot generalizable incremental learning for vision-language object detection.

MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM Zero-shot generalizable incremental learning for vision-language object detection

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-07T12:33:06.634033Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:33:06.634033Z digest=sha256:902206ca5ced741faa7d5416995a4cb675f97353b40e4d8dee6fc324c7062e1a

Observation d23fb554-b6b2-4c64-80be-3802ef89af64 · outbound

This paper cites MR-GDINO: Efficient Open-World Continual Object Detection.

MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM MR-GDINO: Efficient Open-World Continual Object Detection

Reference 9

Resolution
verified exact
local_arxiv, observed 2026-08-07T12:33:14.676390Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T12:33:06.701427Z digest=sha256:9363c2167d2911c129693dc22f22bcffff7ba3eda234925900bccd3a0e9e2d24

Observation f50f399f-d566-4461-af0e-1f313c9b77a1 · outbound

This paper cites Lpt: Long-tailed prompt tuning for image classification.

MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM Lpt: Long-tailed prompt tuning for image classification

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-07T12:33:06.750153Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:33:06.750153Z digest=sha256:21ccb63d259831fab1491195ba966aa4b3c1fb2ea5bbdb8d7f65a647a0294631

Observation 40226a05-9e99-486a-8a25-e68cf003dc0d · outbound

This paper cites A Survey on In-context Learning.

MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM A Survey on In-context Learning

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-07T12:33:06.816254Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:33:06.816254Z digest=sha256:4344698351a4491aaefeefcc8480aa4b61060b434f3dcefcf4539e36c24a9be7

Observation ea35cb9c-7fd7-4ce0-bc0d-e7d524fc5b9f · outbound

This paper cites Virgo: A Preliminary Exploration on Reproducing o1-like MLLM.

MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM Virgo: A Preliminary Exploration on Reproducing o1-like MLLM

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-07T12:33:06.915732Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:33:06.915732Z digest=sha256:0bf4500cb74e9ce991cc06b844f98377d234104c37a007544b7718bbf591d039

Observation 49e1bd80-b7be-43fc-863d-fb2bf75fdb27 · outbound

This paper cites Detecting hallucinations in large language models using semantic entropy.Nature, 630(8017):625–630, 2024.

MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM Detecting hallucinations in large language models using semantic entropy.Nature, 630(8017):625–630, 2024

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-07T12:33:07.022394Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:33:07.022394Z digest=sha256:e61d37992ce7cdb181543dbbc8038be3782b0340eb9b35f742301979d74404a9

Observation b3dc9fb6-d758-4d73-af6e-464ab485d758 · outbound

This paper cites MME: A Comprehensive Evaluation Benchmark for Multimodal Large Language Models.

MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM MME: A Comprehensive Evaluation Benchmark for Multimodal Large Language Models

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-07T12:33:07.090750Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:33:07.090750Z digest=sha256:f4929ea40bdf65b40271e1ab8367f3c129c84d6cda29730910d269e52bbd7613

Observation 6cd15c42-a470-4dbe-9f0d-e99b8db63c67 · outbound

This paper cites GPTScore: Evaluate as you desire.

MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM GPTScore: Evaluate as you desire

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-07T12:33:07.151002Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:33:07.151002Z digest=sha256:0e02a83ed2b4fdd414ade7dd02024343a8034982dc55553157847af49f441090

Observation 85588250-a0aa-44c6-a5fe-3bd9c163fbff · outbound

This paper cites The capacity for moral self-correction in large language models.Parameters, 109(1010):1011.

MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM The capacity for moral self-correction in large language models.Parameters, 109(1010):1011

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-07T12:33:07.271503Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:33:07.271503Z digest=sha256:961cebc20e8ef4a2985ceaf4203bb3bd7bf24431873c58a189f54a0f8b383ac7

Observation 131f8f24-faff-4e6e-9cc9-ecdc554449ba · outbound

This paper cites Are Language Models Puzzle Prodigies? Algorithmic Puzzles Unveil Serious Challenges in Multimodal Reasoning.

MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM Are Language Models Puzzle Prodigies? Algorithmic Puzzles Unveil Serious Challenges in Multimodal Reasoning

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-07T12:33:07.360301Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:33:07.360301Z digest=sha256:75371ab7807eda99bac512ada8d48cb541c6fa47ff4d57d7c90039dc6cbf99db

Observation d385e19b-f4e6-4955-89e5-25a3370bbfea · outbound

This paper cites The Llama 3 Herd of Models.

MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM The Llama 3 Herd of Models

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-07T12:33:07.475863Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:33:07.475863Z digest=sha256:0b38f9eb23021ce7593d29ca965b308d5b1fc30f7a64a4ccc0cdb3bdde08f22c

Observation 82fb3068-1a9f-4e79-a8f4-edb963f00afd · outbound

This paper cites A Survey on LLM-as-a-Judge.

MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM A Survey on LLM-as-a-Judge

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-07T12:33:07.571919Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:33:07.571919Z digest=sha256:b0684dcdad4b27c52c734b11e83b512096efa4e3bf6722729b4e06897404badd

Observation a4bcd11f-f250-480f-bf44-546db8c7a328 · outbound

This paper cites Hallusionbench: an advanced diagnostic suite for entangled language hallucination and visual illusion in large vision-language models.

MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM Hallusionbench: an advanced diagnostic suite for entangled language hallucination and visual illusion in large vision-language models

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-07T12:33:07.664840Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:33:07.664840Z digest=sha256:5a08193bafdbd3d61f581a34ace1bbc52130e492d722b76c9f4f01a480e327f7

Observation aaf1ac63-d878-4cb2-807b-c8f4228efe91 · outbound

This paper cites DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning.

MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-07T12:33:07.772797Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:33:07.772797Z digest=sha256:3e8025ce61aeeec30b7fadd8055770654150a2e2556d5a0faf4c87e00a569863

Observation e4ea4dc8-3afa-4b0e-a85f-889900e9f05d · outbound

This paper cites When Continue Learning Meets Multimodal Large Language Model: A Survey.

MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM When Continue Learning Meets Multimodal Large Language Model: A Survey

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-07T12:33:07.869192Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:33:07.869192Z digest=sha256:edfab16a6291efb4a90bb9f662fdf0c9277ac89067b38e24759141a28bbf14b5

Observation 3c11f0aa-8b18-422f-9198-fd055ab6103c · outbound

This paper cites GPT-4o System Card.

MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM GPT-4o System Card

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-07T12:33:07.959565Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:33:07.959565Z digest=sha256:ade471e250ccdfd629bc4c9825f24640e9a500930148fd4083608d308b4f52b5

Observation efeed5f0-0c9e-4944-a8b8-e7aa228c618e · outbound

This paper cites OpenAI o1 System Card.

MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM OpenAI o1 System Card

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-07T12:33:08.049492Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:33:08.049492Z digest=sha256:6545405d1ddc73e29b9d0b78ad265bcf8c29ab426c5ecc9e204e8daac9dd9b36

Observation ffa0a01f-ae22-4626-935d-4c149f169e29 · outbound

This paper cites Towards mitigating LLM hallucination via self reflection.

MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM Towards mitigating LLM hallucination via self reflection

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:33:18.503341Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T12:33:08.183950Z digest=sha256:950ceaeeaddf59d512439d80f93018f7edd1b4e3b9594c96f8978b10b020197d

Observation 62c94a89-a059-4823-bb1d-acd48b4486d5 · outbound

This paper cites Hallucination augmented contrastive learning for multimodal large language model.

MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM Hallucination augmented contrastive learning for multimodal large language model

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-07T12:33:08.290558Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:33:08.290558Z digest=sha256:6285110f9d43c25f92900b9326f656abde20a486f7ff53c7c253a24cfbf7af61

Observation a72d196d-62cc-4209-83be-37aec9055dc7 · outbound

This paper cites MME-CoT: Benchmarking Chain-of-Thought in Large Multimodal Models for Reasoning Quality, Robustness, and Efficiency.

MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM MME-CoT: Benchmarking Chain-of-Thought in Large Multimodal Models for Reasoning Quality, Robustness, and Efficiency

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-07T12:33:08.379110Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:33:08.379110Z digest=sha256:521b0dc69ddcfb88c2aa4e3cade026a84f47456ecaa89b612c763145fa830a48

Observation d3eabd59-e4da-4a04-a2a9-c88aed94f37e · outbound

This paper cites Decoupling representation and classifier for long-tailed recognition.

MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM Decoupling representation and classifier for long-tailed recognition

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:33:18.166965Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T12:33:08.480767Z digest=sha256:d78df90473c9168013122adda976bed8ba2a20f1694281493ea8c366708169c0

Observation 2e8ba4f8-f945-4fb9-9e87-57bb85616c2f · outbound

This paper cites Gonzalez, Hao Zhang, and Ion Stoica.

MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM Gonzalez, Hao Zhang, and Ion Stoica

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-07T12:33:08.559186Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:33:08.559186Z digest=sha256:2515a1a716fa3ea490da509d7b22207e73466433ae7900ecff022d53b72bbb71

Observation c3e2ac17-0813-463f-9f74-946b2dfbbd0c · outbound

This paper cites SEED-Bench: Benchmarking Multimodal LLMs with Generative Comprehension.

MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM SEED-Bench: Benchmarking Multimodal LLMs with Generative Comprehension

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-07T12:33:08.631430Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:33:08.631430Z digest=sha256:633941a13df67587a43a8123bebbd6fb2cb58531582940ed5a1ef299fa37a9dd

Observation 5dc004e3-889e-47bb-aa86-7f110d4a3be6 · outbound

This paper cites LLMs-as-Judges: A Comprehensive Survey on LLM-based Evaluation Methods.

MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM LLMs-as-Judges: A Comprehensive Survey on LLM-based Evaluation Methods

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-07T12:33:08.729239Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:33:08.729239Z digest=sha256:df2e6d7ff9e223921caa30bd1c9a19cf4e6dd09ce753fd1b2bcd2c40d7d6e743

Observation 3042c996-c970-4899-9c34-7b2afc9e70d9 · outbound

This paper cites Salad-bench: A hierarchical and comprehensive safety benchmark for large language models.

MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM Salad-bench: A hierarchical and comprehensive safety benchmark for large language models

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-07T12:33:08.797773Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:33:08.797773Z digest=sha256:95dfc0e293e95f72a8099b6e1a14c4c90a7c06677ad68e491a457a22d440db50

Observation ebe72b6c-d1fb-434d-b5a8-7dd39b936bf4 · outbound

This paper cites Long-tailed visual recognition via gaussian clouded logit adjustment.

MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM Long-tailed visual recognition via gaussian clouded logit adjustment

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:33:17.903767Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T12:33:08.934739Z digest=sha256:537b82845a7739b469f9779cd5b80dd382b9617a705514b14b48cb6d863f1c1d

Observation 4b24e925-a8de-49f7-ae0e-ae79231d2290 · outbound

This paper cites Evaluating object hallucination in large vision-language models.

MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM Evaluating object hallucination in large vision-language models

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:33:17.659774Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T12:33:09.006807Z digest=sha256:e71fb4dd633a63b6472154e048ff8227aa6f1e77ab3320da29aa930439c20184

Observation a46d00ea-67a9-4727-8ff8-6fe63c8f0dda · outbound

This paper cites Omnibench: Towards the future of universal omni-language models.arXiv preprint arXiv:2409.15272, 2024.

MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM Omnibench: Towards the future of universal omni-language models.arXiv preprint arXiv:2409.15272, 2024

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-07T12:33:09.096630Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:33:09.096630Z digest=sha256:a5187cadc5ce39373d79b4e930244f153f0cc3b37e41ee6a1fef8454d52a90f1

Observation ef6c1503-b847-4245-861e-fa84899b4c71 · outbound

This paper cites DeepSeek-V3 Technical Report.

MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM DeepSeek-V3 Technical Report

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-07T12:33:09.169761Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:33:09.169761Z digest=sha256:3c0a6260ef4ef70290e1acfa594b82edc5f1c8982e4718ee7fb1c1fdd01ac966

Observation 169dedb4-ccf8-4215-b20e-99bfc5d5e067 · outbound

This paper cites Mitigat- ing hallucination in large multi-modal models via robust instruction tuning.

MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM Mitigat- ing hallucination in large multi-modal models via robust instruction tuning

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:33:17.441938Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T12:33:09.244473Z digest=sha256:fcd6fdb6de60419177bc8f64547d5aa9b1947c602b858600ea2f84a771757144

Observation a549b860-d5f8-4579-88d9-6e0627e2fb8f · outbound

This paper cites Visual instruction tuning, 2023.

MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM Visual instruction tuning, 2023

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:33:17.186401Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T12:33:09.337797Z digest=sha256:941b668dab60469e31e3f823cb4e1aeaf2e2665b011c3aad41b9d59d852d4612

Observation 2551a763-37bd-41ce-931b-e4db63c3bf7d · outbound

This paper cites Visual-RFT: Visual Reinforcement Fine-Tuning.

MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM Visual-RFT: Visual Reinforcement Fine-Tuning

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-07T12:33:09.432965Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:33:09.432965Z digest=sha256:f1eab89b02120d5207947033cfd0efcde57919694ed3886f13e83513f74e0e81

Observation 7f6f5d91-2270-4ad3-be3f-e631d7dceb6c · outbound

This paper cites Decoupled weight decay regularization.

MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM Decoupled weight decay regularization

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-07T12:33:09.543779Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:33:09.543779Z digest=sha256:1733e887935f8f8bfdfbbb0bbbde5037b2e1caa76506d36f55c8b3f154051a08

Observation b7779aaa-102d-49ed-8517-7739ee88a52b · outbound

This paper cites Mathvista: Evaluating mathematical reasoning of foundation models in visual contexts.

MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM Mathvista: Evaluating mathematical reasoning of foundation models in visual contexts

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-07T12:33:09.643603Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:33:09.643603Z digest=sha256:dbce9a47a28c0cfbf72cc8083f7069496c0f4178dde32450a2965320869175a0

Observation 50b2646c-45b8-47df-92bf-d05cb8e0947f · outbound

This paper cites Learn to explain: Multimodal reasoning via thought chains for science question answering.

MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM Learn to explain: Multimodal reasoning via thought chains for science question answering

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-07T12:33:09.735339Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:33:09.735339Z digest=sha256:8424f96f8455c0f6cb6676c65336f310f877309d5f0db339c8a1d455bb381472

Observation 0c0d71be-894d-4563-aa5a-2caaf14733de · outbound

This paper cites Ursa: Understanding and verifying chain-of-thought reasoning in multimodal mathematics.arXiv preprint arXiv:2501.04686, 2025.

MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM Ursa: Understanding and verifying chain-of-thought reasoning in multimodal mathematics.arXiv preprint arXiv:2501.04686, 2025

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-07T12:33:09.808767Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:33:09.808767Z digest=sha256:e8621bcf6c052926ef3578950a05e74c602ac7dc7fb7833bc8e9be94d0d4601f

Observation 63567212-d451-45be-8661-10868da2d31b · outbound

This paper cites Ok-vqa: A visual question answering benchmark requiring external knowledge.

MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM Ok-vqa: A visual question answering benchmark requiring external knowledge

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-07T12:33:09.913245Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:33:09.913245Z digest=sha256:dbc05b186dc0fa381342f4671cad91469267218b2f38b48a344b870a7467bb78

Observation d1fc6cd4-025a-416c-a8d9-0a324dbcea08 · outbound

This paper cites Chartqa: A benchmark for question answering about charts with visual and logical reasoning.

MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM Chartqa: A benchmark for question answering about charts with visual and logical reasoning

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-07T12:33:09.990090Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:33:09.990090Z digest=sha256:33f8525b8b2de5e1da82da7fbaf70f11b33fdb4f522a61b7fd7c51ee9eee9be4

Observation d12ecdd1-663f-4408-a3f2-485ff7baeab6 · outbound

This paper cites MM-Eureka: Exploring the Frontiers of Multimodal Reasoning with Rule-based Reinforcement Learning.

MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM MM-Eureka: Exploring the Frontiers of Multimodal Reasoning with Rule-based Reinforcement Learning

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-07T12:33:10.071390Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:33:10.071390Z digest=sha256:9624e3b466b26c3e8911c617e867420b07017e2e0e00b86f5ec16ec9bbb9ccbe

Observation 4d684078-fce3-4628-b438-3871fe71aeab · outbound

This paper cites Compositional chain-of- thought prompting for large multimodal models.

MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM Compositional chain-of- thought prompting for large multimodal models

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-07T12:33:10.139052Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:33:10.139052Z digest=sha256:be304f316ec2893310e231d3e0e540a7221c2a716d1bba961a0bd714b3a8c132

Observation a53719f6-9c58-428f-b724-d963b96de839 · outbound

This paper cites Pytorch: An imperative style, high-performance deep learning library.

MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM Pytorch: An imperative style, high-performance deep learning library

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-07T12:33:10.217648Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:33:10.217648Z digest=sha256:7f8101c2410e615b806e76f4b1272d7077abf27c8503e4e838e3a467bd21b742

Observation 04a56abd-815d-43aa-81bf-37bc82357b95 · outbound

This paper cites LMM-R1: Empowering 3B LMMs with Strong Reasoning Abilities Through Two-Stage Rule-Based RL.

MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM LMM-R1: Empowering 3B LMMs with Strong Reasoning Abilities Through Two-Stage Rule-Based RL

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-07T12:33:10.310457Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:33:10.310457Z digest=sha256:9a885ddf5fc229f1070517765f00bb4d30aeb2248b617b52daf4a52908d08336

Observation 9e83b7d7-971b-4220-8f1f-5ce5d603d714 · outbound

This paper cites ReCEval: Evaluating Reasoning Chains via Correctness and Informativeness.

MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM ReCEval: Evaluating Reasoning Chains via Correctness and Informativeness

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-07T12:33:10.405095Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:33:10.405095Z digest=sha256:2d46c1ea300f68dcd5b992d4750ebc8d72a6bba84ccb6e28f6542e07417c035d

Observation 698b7e53-57d5-4cb9-a2ee-331fa319c673 · outbound

This paper cites Mitigating object hallucination in mllms via data-augmented phrase-level alignment.

MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM Mitigating object hallucination in mllms via data-augmented phrase-level alignment

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:33:16.787632Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T12:33:10.547565Z digest=sha256:86cd0bc498c516c52bbc276095cf93ecf04c8afd949d106196c9bca5c2b6f865

Observation 284f205c-e24d-49d4-98ed-c7382e884264 · outbound

This paper cites Proximal Policy Optimization Algorithms.

MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM Proximal Policy Optimization Algorithms

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-07T12:33:10.633888Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:33:10.633888Z digest=sha256:a5d6ed72d8d54c3c7b5de7f38470617c582bf08c1304eadf1622da52c30a8fb9

Observation a0bec2fa-ad09-47da-9b3d-dcfc443a20cb · outbound

This paper cites DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models.

MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-07T12:33:10.745345Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:33:10.745345Z digest=sha256:af966525defd19b7b537234f2dde4854ee1a4bd16a2fb89405f3c86588cc151e

Observation b2b2aefe-d4c4-4b50-9c4c-dffc83574f91 · outbound

This paper cites Monte carlo tree search: A review of recent modifications and applications.Artificial Intelligence Review, 56(3):2497–2562, 2023.

MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM Monte carlo tree search: A review of recent modifications and applications.Artificial Intelligence Review, 56(3):2497–2562, 2023

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-07T12:33:10.856167Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:33:10.856167Z digest=sha256:bf0dd75466ff1b9b921571b222b59cc9cd1e77640fd2123bd96bb0027cc59329

Observation dcba0640-8e74-499b-936a-ba7fdccf62b7 · outbound

This paper cites Gemini: A Family of Highly Capable Multimodal Models.

MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM Gemini: A Family of Highly Capable Multimodal Models

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-07T12:33:10.924496Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:33:10.924496Z digest=sha256:48384f76fd4b1e6885fe07f1fd714f611e5de23e440a29b815cca5a95cb7c6cb

Observation 443738d3-a1a4-4e01-8fa7-10e641d3eecb · outbound

This paper cites QVQ: To See the World with Wisdom.

MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM QVQ: To See the World with Wisdom

Reference 56

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:33:16.606887Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T12:33:11.048068Z digest=sha256:45f3721b50d537be71aad1a0a01829ecb08a1a838ac46f9007a8097aadb464be

Observation 61e93b22-71a1-4bf0-957f-2c3634b8c1f2 · outbound

This paper cites Uncertainty-Based Abstention in LLMs Improves Safety and Reduces Hallucinations.

MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM Uncertainty-Based Abstention in LLMs Improves Safety and Reduces Hallucinations

Reference 57

Resolution
unresolved
no resolver link, observed 2026-08-07T12:33:11.153366Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:33:11.153366Z digest=sha256:c4977379d3aae760171d0216264d47acdb7ff8d773c3fc561abb84c68ecb899d

Observation 8084a308-e4ac-4805-9d83-a724b7daae11 · outbound

This paper cites Eyes wide shut? exploring the visual shortcomings of multimodal llms.

MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM Eyes wide shut? exploring the visual shortcomings of multimodal llms

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-07T12:33:11.217618Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:33:11.217618Z digest=sha256:54ff2fe14a9614aba40b4914bddcff7aaf4aef935756da7282d9cf512d15bd7b

Observation e08b3ec0-020d-434d-8a05-7a02209f9913 · outbound

This paper cites Measuring multimodal mathematical reasoning with math-vision dataset.

MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM Measuring multimodal mathematical reasoning with math-vision dataset

Reference 59

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:33:16.486557Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T12:33:11.281006Z digest=sha256:e17f912b330c50cb1ea190e54d13b37fe7c2f9a8c73b8c66ef5b23b104b7ff0b

Observation 22da4bca-b3c8-4658-8779-f635c78267e6 · outbound

This paper cites Qwen2-VL: Enhancing Vision-Language Model's Perception of the World at Any Resolution.

MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM Qwen2-VL: Enhancing Vision-Language Model's Perception of the World at Any Resolution

Reference 60

Resolution
unresolved
no resolver link, observed 2026-08-07T12:33:11.350592Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:33:11.350592Z digest=sha256:e4d7d0609dd7200c375e5e5ba3c7245a5f0685036452aeb391e0136f2cfac020

Observation 1772c607-1e3e-4cc8-a011-797bd52d4277 · outbound

This paper cites Chain-of-thought prompting elicits reasoning in large language models.

MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM Chain-of-thought prompting elicits reasoning in large language models

Reference 61

Resolution
unresolved
no resolver link, observed 2026-08-07T12:33:11.422440Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:33:11.422440Z digest=sha256:ab5ec78890e566e1393e4d80af0a072213bcc2e323beeb5873447953e146953c

Observation edbb8b53-7b56-4357-9147-0fd3e8b09e71 · outbound

This paper cites RITUAL: Random Image Transformations as a Universal Anti-hallucination Lever in Large Vision Language Models.

MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM RITUAL: Random Image Transformations as a Universal Anti-hallucination Lever in Large Vision Language Models

Reference 62

Resolution
unresolved
no resolver link, observed 2026-08-07T12:33:11.492919Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:33:11.492919Z digest=sha256:e75396c6d8ba3e73dec7f4d674a7d4ef0bfca8bc518e177410a408b218c27f34

Observation 630b0b44-60ff-4031-b5bd-136c5134da95 · outbound

This paper cites Autohallusion: Automatic gen- eration of hallucination benchmarks for vision-language models.

MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM Autohallusion: Automatic gen- eration of hallucination benchmarks for vision-language models

Reference 63

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:33:16.317333Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T12:33:11.562584Z digest=sha256:e83012c4fb3816652775c017191da5f71dc3a1983e041e4c62939c3a0796a911

Observation 2767d4b3-7a1e-44f6-9233-16d9eaf5c5ab · outbound

This paper cites Grok 3 Beta — The Age of Reasoning Agents.https://x.ai/grok, 2025.

MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM Grok 3 Beta — The Age of Reasoning Agents.https://x.ai/grok, 2025

Reference 64

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:33:16.123781Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T12:33:11.613404Z digest=sha256:e762b2419c14af61107377e8df634d94f849e84f326bdc4e1f5503ac1b1d3f4a

Observation 9b1c93ed-31f5-4c5e-aef0-419a9196b660 · outbound

This paper cites Mitigating object hallucination via concentric causal attention.Advances in Neural Information Processing Systems, 37:92012–92035, 2024.

MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM Mitigating object hallucination via concentric causal attention.Advances in Neural Information Processing Systems, 37:92012–92035, 2024

Reference 65

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:33:15.957279Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T12:33:11.724385Z digest=sha256:d1250e0b123a1e085487721b6792a95fa7b89e2acd10a1318d02c772ebe61a03

Observation 6ce3eff8-0eda-42fa-bd40-3d76fa9bab4f · outbound

This paper cites Llava-cot: Let vision language models reason step-by-step, 2024.

MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM Llava-cot: Let vision language models reason step-by-step, 2024

Reference 66

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:33:15.820797Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T12:33:11.795732Z digest=sha256:8616f3fd2cd3fe78dbaa1f45d14b8c4544f6f263be10c9e90aab1656b62a6f3b

Observation eebbacd3-290d-40b6-9616-cec260b62609 · outbound

This paper cites Qwen2.5 Technical Report.

MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM Qwen2.5 Technical Report

Reference 67

Resolution
unresolved
no resolver link, observed 2026-08-07T12:33:11.894050Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:33:11.894050Z digest=sha256:e14a71cb01b0382f50b03176c763fe74d762fa7303acf60e3b40304de86e8181

Observation 98b48a0d-4c30-470a-b2d7-380e837c3767 · outbound

This paper cites Soft-prompting with graph-of-thought for multi-modal representation learning.

MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM Soft-prompting with graph-of-thought for multi-modal representation learning

Reference 68

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:33:15.698313Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T12:33:11.952400Z digest=sha256:6797df99fd4b3fa5f4a26c3e7e260cd7f9e773339eea20d4290a3c9bec4154c3

Observation 9304223c-8c7d-4574-b644-7bb3046630c9 · outbound

This paper cites Mitigating hallucination in large vision- language models via modular attribution and intervention.

MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM Mitigating hallucination in large vision- language models via modular attribution and intervention

Reference 69

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:33:15.543195Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T12:33:12.053954Z digest=sha256:e77d2f10c23c3a29f2111149d5434aef05e4041d1cb81e12512c90bc845be4c6

Observation 73dcefd7-a958-4dd6-bded-72c5e627c82b · outbound

This paper cites R1-Onevision: Advancing Generalized Multimodal Reasoning through Cross-Modal Formalization.

MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM R1-Onevision: Advancing Generalized Multimodal Reasoning through Cross-Modal Formalization

Reference 70

Resolution
unresolved
no resolver link, observed 2026-08-07T12:33:12.177383Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:33:12.177383Z digest=sha256:f16ac2b36cdbeaaa75afcd11d2db81d77857e1cf786ad2d9d145896aca980042

Observation 0b48fbce-4c79-415a-aeac-4e3e872724ae · outbound

This paper cites Mulberry: Empowering MLLM with o1-like Reasoning and Reflection via Collective Monte Carlo Tree Search.

MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM Mulberry: Empowering MLLM with o1-like Reasoning and Reflection via Collective Monte Carlo Tree Search

Reference 71

Resolution
unresolved
no resolver link, observed 2026-08-07T12:33:12.258432Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:33:12.258432Z digest=sha256:87d21ab91e049a8ffff67abea1b6fab2e6b7d8bb0ebe72b16549d3b43bc238bb

Observation c618bb1f-9378-4f4e-a8e8-05ca272c661b · outbound

This paper cites React: Synergizing reasoning and acting in language models.

MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM React: Synergizing reasoning and acting in language models

Reference 72

Resolution
unresolved
no resolver link, observed 2026-08-07T12:33:12.383675Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:33:12.383675Z digest=sha256:1d1ebf0291948e7af62c6161786a3a96a788aaa33da5c121d01fcb2cc49e0958

Observation ed50d571-49f4-4efd-a3c0-5fc29315c7a7 · outbound

This paper cites LIMO: Less is More for Reasoning.

MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM LIMO: Less is More for Reasoning

Reference 73

Resolution
unresolved
no resolver link, observed 2026-08-07T12:33:12.448500Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:33:12.448500Z digest=sha256:73a210c3c240146391e9355292fa2e22be5af3ea5746862c775574227b6db2a2

Observation 3cb4a8e8-4452-4317-ac9e-074e4f911ef1 · outbound

This paper cites Hallucidoctor: Mitigating hallucinatory toxicity in visual instruction data.

MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM Hallucidoctor: Mitigating hallucinatory toxicity in visual instruction data

Reference 74

Resolution
unresolved
no resolver link, observed 2026-08-07T12:33:12.557605Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:33:12.557605Z digest=sha256:3cc8dcfaa4f97209b8266136fe7160c3b702031c0f4e6fefae7c63236fe2f3fe

Observation 9d8616b4-cc9f-478c-83fe-fa1433f61ea9 · outbound

This paper cites DAPO: An Open-Source LLM Reinforcement Learning System at Scale.

MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM DAPO: An Open-Source LLM Reinforcement Learning System at Scale

Reference 75

Resolution
unresolved
no resolver link, observed 2026-08-07T12:33:12.653152Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:33:12.653152Z digest=sha256:a5df23b8566feb7002339529ae574724287f6f058846bf1c79473c19ca01e457

Observation 0f2fb295-f912-4fbd-8059-15d7e34eb919 · outbound

This paper cites Rlhf-v: Towards trustworthy mllms via behavior alignment from fine-grained correctional human feedback.

MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM Rlhf-v: Towards trustworthy mllms via behavior alignment from fine-grained correctional human feedback

Reference 76

Resolution
unresolved
no resolver link, observed 2026-08-07T12:33:12.750492Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:33:12.750492Z digest=sha256:67c6d09fa4defb69e4040b905776ddb182c28fad5d1b470fb75fcfec6aefa80c

Observation 9ebaadc9-87c8-41a0-ba69-f23f61aa0c81 · outbound

This paper cites LLaMA-Berry: Pairwise Optimization for O1-like Olympiad-Level Mathematical Reasoning.

MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM LLaMA-Berry: Pairwise Optimization for O1-like Olympiad-Level Mathematical Reasoning

Reference 77

Resolution
unresolved
no resolver link, observed 2026-08-07T12:33:12.857251Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:33:12.857251Z digest=sha256:be96bffcc1b77b9cbae9728b62dfa4a0420b6864463252ab120c86f32a0b09ac

Observation 5e0dc656-f057-48a7-a72b-70435f63512d · outbound

This paper cites Reflective instruction tuning: Mitigating hallucinations in large vision-language models.

MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM Reflective instruction tuning: Mitigating hallucinations in large vision-language models

Reference 78

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:33:15.380205Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T12:33:12.980715Z digest=sha256:3676f44db0cf5b2749a7c8525a96cf2a22701ed31d5e26b36806e95ba3953eaa

Observation 5a298254-9648-4782-abb3-752e69622b8f · outbound

This paper cites Mathverse: Does your multi-modal llm truly see the diagrams in visual math problems? InEuropean Conference on Computer Vision, pages 169–186.

MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM Mathverse: Does your multi-modal llm truly see the diagrams in visual math problems? InEuropean Conference on Computer Vision, pages 169–186

Reference 79

Resolution
unresolved
no resolver link, observed 2026-08-07T12:33:13.073898Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:33:13.073898Z digest=sha256:853d565c9dd632239e0a66116cd6ea813d54e09a3bac5c9843d9f30be565f7d0

Observation 140db751-7cb8-4809-9e30-2e832334feb1 · outbound

This paper cites VL-Uncertainty: Detecting Hallucination in Large Vision-Language Model via Uncertainty Estimation.

MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM VL-Uncertainty: Detecting Hallucination in Large Vision-Language Model via Uncertainty Estimation

Reference 80

Resolution
unresolved
no resolver link, observed 2026-08-07T12:33:13.190579Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:33:13.190579Z digest=sha256:a2acdd8f8a6fc85a3ddf32c3ec8e7050a1942c5f9bab904e90f256c0823584cb

Observation 334b295f-408d-44bd-8026-aa20d981a015 · outbound

This paper cites The Lessons of Developing Process Reward Models in Mathematical Reasoning.

MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM The Lessons of Developing Process Reward Models in Mathematical Reasoning

Reference 81

Resolution
unresolved
no resolver link, observed 2026-08-07T12:33:13.281932Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:33:13.281932Z digest=sha256:10bde2609e3be5dd1e2eb6f78cbb2176e31c003345e3885ef19285db11cd6ae6

Observation 5fd38f61-18d5-478f-b822-1218dce4fd63 · outbound

This paper cites Multimodal chain-of-thought reasoning in language models.Transactions on Machine Learning Research.

MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM Multimodal chain-of-thought reasoning in language models.Transactions on Machine Learning Research

Reference 82

Resolution
unresolved
no resolver link, observed 2026-08-07T12:33:13.378434Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:33:13.378434Z digest=sha256:293ddc55f7e68f29d3065a1f192936cc54b0fb22db0529e224b85b736d41eb9a

Observation daedcae0-b3a7-437e-b406-b2b3c87dc327 · outbound

This paper cites Beyond Hallucinations: Enhancing LVLMs through Hallucination-Aware Direct Preference Optimization.

MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM Beyond Hallucinations: Enhancing LVLMs through Hallucination-Aware Direct Preference Optimization

Reference 83

Resolution
unresolved
no resolver link, observed 2026-08-07T12:33:13.470889Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:33:13.470889Z digest=sha256:943d97d654a257a15556a8e00694a1bf11364e61d6d00ed6a474e351550fd6d4

Observation cfacc5c1-dfc2-4639-8b13-4af198dd8327 · outbound

This paper cites A picture is worth a graph: A blueprint debate paradigm for multimodal reasoning.

MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM A picture is worth a graph: A blueprint debate paradigm for multimodal reasoning

Reference 84

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:33:15.196268Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T12:33:13.563966Z digest=sha256:09f1f8866430233376dd5cc9253cae3c60d93ad97a1206a30048d39386185842

Observation 7f35f66e-4cbe-463e-a511-e8f97f085ad1 · outbound

This paper cites Thinking Before Looking: Improving Multimodal LLM Reasoning via Mitigating Visual Hallucination.

MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM Thinking Before Looking: Improving Multimodal LLM Reasoning via Mitigating Visual Hallucination

Reference 85

Resolution
unresolved
no resolver link, observed 2026-08-07T12:33:13.663322Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:33:13.663322Z digest=sha256:599689b82b5a019d7c6feb9efd2683457b2d2d9712b43f575914e7edbf7dae27

Observation 0db258ac-a6f2-4765-b849-f5e6ecfc0fd7 · outbound

This paper cites Relying on the unreliable: The impact of language models’ reluctance to express uncertainty.

MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM Relying on the unreliable: The impact of language models’ reluctance to express uncertainty

Reference 86

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:33:15.038180Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T12:33:13.747662Z digest=sha256:d9ad0cbc6ff529399e17dd46c5e3a65976db27d5f64d7b562dc00821043d8167

Observation 0d179a1e-f3fd-41b8-a842-23e070afc6e7 · outbound

This paper cites the answer is [answer in the input].

MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM the answer is [answer in the input]

Reference 87

Resolution
malformed identifier
raw_fallback, observed 2026-08-07T12:33:14.881658Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T12:33:13.820971Z digest=sha256:a48ced761dc344e73b7095775a3ed6cc4c17a63a2a7583ac387eb0503c551baa

Pith citing papers

Observation 8f61cce6-5c2d-4d17-8b3f-2c8cc3ad9045 · inbound

A Comprehensive Survey on Trustworthiness in Reasoning with Large Language Models cites this paper.

A Comprehensive Survey on Trustworthiness in Reasoning with Large Language Models MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM

Reference 62

Resolution
unresolved
no resolver link, observed 2026-08-05T10:39:03.028373Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:39:03.028373Z digest=sha256:bcb0a91ce799688cca185a46286e0a1c83d35baa864249ec028ab49ca9bb50f6

Observation 93d009a1-1467-4909-ae1a-aa45d4e09569 · inbound

Understanding the Role of Hallucination in Reinforcement Post-Training of Multimodal Reasoning Models cites this paper.

Understanding the Role of Hallucination in Reinforcement Post-Training of Multimodal Reasoning Models MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM

Reference 6

Resolution
verified exact
arxiv_id, observed 2026-05-13T20:53:16.282414Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-13T20:48:52.130130Z digest=sha256:b31b5ae0a5bab91e7b320517f9da74a95c13eea8e5f934946a743dc73fdeae31

Observation 8db78eb4-7cbe-4125-83e5-637815967363 · inbound

HypEHR: Hyperbolic Modeling of Electronic Health Records for Efficient Question Answering cites this paper.

HypEHR: Hyperbolic Modeling of Electronic Health Records for Efficient Question Answering MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM

Reference 230

Resolution
metadata mismatch
arxiv_id, observed 2026-05-09T23:54:45.651801Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-05-09T23:51:47.724033Z digest=sha256:3b598e007653a5b7ba815470652cc2457ee12d0b15e9de7959793d9f5637a3e5

Observation 45eb3f94-6839-4143-946f-13d83c6c4310 · inbound

SoccerRef-Agents: Multi-Agent System for Automated Soccer Refereeing cites this paper.

SoccerRef-Agents: Multi-Agent System for Automated Soccer Refereeing MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM

Reference 5

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T20:46:13.723544Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-08T08:03:57.287297Z digest=sha256:6670c832b4b4e9dc85a30c568590e8966a513419913e587afb8d8f054f45188c