Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-12T19:35:45.274312Z
Paper Citation Record · LEDGER
As of 21 August 2026, this Paper Citation Record lists 49 of 49 outbound references and 10 inbound Pith citation observations for arXiv:2411.12591.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-12T19:35:45.274312Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-20T06:33:59.587034+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-15T17:10:02.962883Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-05-15T17:18:53.747034Z
49 of 49 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 3e6706d4-3415-4aac-a73f-798fa6a258f8 · outbound
Thinking Before Looking: Improving Multimodal LLM Reasoning via Mitigating Visual Hallucination GPT-4 Technical Report
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fccdf2c1-e8fd-4853-8af9-18c9613a300f · outbound
Thinking Before Looking: Improving Multimodal LLM Reasoning via Mitigating Visual Hallucination Qwen-VL: A Versatile Vision-Language Model for Understanding, Localization, Text Reading, and Beyond
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9df722e4-4254-40c7-8c40-9e6bd7845bdd · outbound
Thinking Before Looking: Improving Multimodal LLM Reasoning via Mitigating Visual Hallucination Hallucination of Multimodal Large Language Models: A Survey
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 418e0488-da97-4b86-837a-c4819cae40e6 · outbound
Thinking Before Looking: Improving Multimodal LLM Reasoning via Mitigating Visual Hallucination Visual question answering on image sets
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation d3a42c52-c14f-4e79-8351-059360ab9996 · outbound
Thinking Before Looking: Improving Multimodal LLM Reasoning via Mitigating Visual Hallucination Rubi: Reducing unimodal biases for visual question answering
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation e2bc77d8-2610-458f-b0f9-44570dff61e6 · outbound
Thinking Before Looking: Improving Multimodal LLM Reasoning via Mitigating Visual Hallucination Gemini 1.5: Unlocking multimodal un- derstanding across millions of tokens of context
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 003119e3-3d0c-488e-9647-a969312f1d80 · outbound
Thinking Before Looking: Improving Multimodal LLM Reasoning via Mitigating Visual Hallucination Learning to prompt for open-vocabulary ob- ject detection with vision-language model
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0dddba74-901b-4687-bc9e-05c453ba19d6 · outbound
Thinking Before Looking: Improving Multimodal LLM Reasoning via Mitigating Visual Hallucination Mme: A compre- hensive evaluation benchmark for multimodal large language models, 2024
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 3a467f93-04ca-4118-8162-befb31904b1a · outbound
Thinking Before Looking: Improving Multimodal LLM Reasoning via Mitigating Visual Hallucination Complexity-based prompting for multi-step reasoning
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 3ea81aaf-bce2-4f73-8bf6-94b8b0f89cb6 · outbound
Thinking Before Looking: Improving Multimodal LLM Reasoning via Mitigating Visual Hallucination Hallusionbench: an advanced diagnos- tic suite for entangled language hallucination and visual il- lusion in large vision-language models
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation fe1b2a44-f5cc-431c-bf8c-c1c1ae72c523 · outbound
Thinking Before Looking: Improving Multimodal LLM Reasoning via Mitigating Visual Hallucination Reasoning with language model is planning with world model, 2023
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation b710a177-c817-44eb-acd7-d705215ae4f0 · outbound
Thinking Before Looking: Improving Multimodal LLM Reasoning via Mitigating Visual Hallucination Towards reason- ing in large language models: A survey
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 7605b9e2-e2cf-4856-8add-bf8b6aba6039 · outbound
Thinking Before Looking: Improving Multimodal LLM Reasoning via Mitigating Visual Hallucination FaithScore: Fine-grained Evaluations of Hallucinations in Large Vision-Language Models
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 48e1db89-2dfe-469c-8d84-36cb4c85c192 · outbound
Thinking Before Looking: Improving Multimodal LLM Reasoning via Mitigating Visual Hallucination Seed-bench: Bench- marking multimodal large language models
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 71a2c00d-b5c0-405a-b7ca-4b598b2ab2be · outbound
Thinking Before Looking: Improving Multimodal LLM Reasoning via Mitigating Visual Hallucination Making Large Language Models Better Reasoners with Step-Aware Verifier
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e8ca54f3-666d-45ec-a171-3ede520f9401 · outbound
Thinking Before Looking: Improving Multimodal LLM Reasoning via Mitigating Visual Hallucination Evaluating Object Hallucination in Large Vision-Language Models
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8ba21dcb-707c-4579-b659-6f34f7ed9736 · outbound
Thinking Before Looking: Improving Multimodal LLM Reasoning via Mitigating Visual Hallucination HallusionBench: An Advanced Diagnostic Suite for Entangled Language Hallucination and Visual Illusion in Large Vision-Language Models
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 47b632c8-6cac-4f46-b534-b63e918fa968 · outbound
Thinking Before Looking: Improving Multimodal LLM Reasoning via Mitigating Visual Hallucination Mitigating Hallucination in Large Multi-Modal Models via Robust Instruction Tuning
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1c6e49f5-f8e6-4e92-ac46-b989b08edcc4 · outbound
Thinking Before Looking: Improving Multimodal LLM Reasoning via Mitigating Visual Hallucination Mitigating hallucination in large multi-modal models via robust instruction tuning
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d4d8b8ae-df91-4644-9f1c-b54758c07750 · outbound
Thinking Before Looking: Improving Multimodal LLM Reasoning via Mitigating Visual Hallucination Improved baselines with visual instruction tuning
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 90cc9f59-b33a-49a9-8b6d-bfe8d149ab19 · outbound
Thinking Before Looking: Improving Multimodal LLM Reasoning via Mitigating Visual Hallucination Visual instruction tuning
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bcc4d45e-8e47-4304-92e6-8782866ebf17 · outbound
Thinking Before Looking: Improving Multimodal LLM Reasoning via Mitigating Visual Hallucination TempCompass: Do Video LLMs Really Understand Videos?
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ed4a26d4-44b4-4ea9-afd2-a0c074532fe5 · outbound
Thinking Before Looking: Improving Multimodal LLM Reasoning via Mitigating Visual Hallucination MathVista: Evaluating Mathematical Reasoning of Foundation Models in Visual Contexts
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4b294586-5ab2-4bbf-8dc7-ac70814d972a · outbound
Thinking Before Looking: Improving Multimodal LLM Reasoning via Mitigating Visual Hallucination Compositional chain-of-thought prompting for large multimodal models
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation c2ab1977-10c7-4792-b7f2-9b8883e4bb6b · outbound
Thinking Before Looking: Improving Multimodal LLM Reasoning via Mitigating Visual Hallucination Kam-cot: Knowledge augmented multimodal chain-of-thoughts reasoning
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 5f0b499f-b78a-4530-8bb4-1132459c84e8 · outbound
Thinking Before Looking: Improving Multimodal LLM Reasoning via Mitigating Visual Hallucination Gpt-4o system card
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation d7611116-14db-4f4e-afc9-72a8cfe821ff · outbound
Thinking Before Looking: Improving Multimodal LLM Reasoning via Mitigating Visual Hallucination Learning transferable visual models from natural language supervi- sion
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b8c425de-b99d-4d42-b227-4a89295de079 · outbound
Thinking Before Looking: Improving Multimodal LLM Reasoning via Mitigating Visual Hallucination Learning To Retrieve Prompts for In-Context Learning
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bfbb3fde-d773-4201-84ec-34ab9fc8bfd5 · outbound
Thinking Before Looking: Improving Multimodal LLM Reasoning via Mitigating Visual Hallucination Visual CoT: Advancing Multi-Modal Language Models with a Comprehensive Dataset and Benchmark for Chain-of-Thought Reasoning
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 651e6ade-92f5-40a6-a53f-b316294644c9 · outbound
Thinking Before Looking: Improving Multimodal LLM Reasoning via Mitigating Visual Hallucination Cognitive psychology
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation ad6ad5e7-c6b1-4451-8e09-448e4aee02c1 · outbound
Thinking Before Looking: Improving Multimodal LLM Reasoning via Mitigating Visual Hallucination Expectation (and attention) in visual cognition
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation e6511a7b-8b8f-4f33-848f-39170cec7c3f · outbound
Thinking Before Looking: Improving Multimodal LLM Reasoning via Mitigating Visual Hallucination EVA-CLIP: Improved Training Techniques for CLIP at Scale
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d6fa9169-ed14-42e0-b2a2-0d24ed339044 · outbound
Thinking Before Looking: Improving Multimodal LLM Reasoning via Mitigating Visual Hallucination Eyes wide shut? exploring the visual shortcomings of multimodal llms
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 7fc8af1f-7def-4fa6-95e6-40274a45617e · outbound
Thinking Before Looking: Improving Multimodal LLM Reasoning via Mitigating Visual Hallucination LLaMA: Open and Efficient Foundation Language Models
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 47b4dfc1-bf0a-4a5a-b01b-aaf1c93cba43 · outbound
Thinking Before Looking: Improving Multimodal LLM Reasoning via Mitigating Visual Hallucination Cross modality bias in visual question answering: A causal view with possible worlds vqa
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 21645259-89ca-4f6f-97d2-526a508de09c · outbound
Thinking Before Looking: Improving Multimodal LLM Reasoning via Mitigating Visual Hallucination Vigc: Visual instruction generation and correction
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 063d4dfd-0678-4759-ab06-fcb52b1123ae · outbound
Thinking Before Looking: Improving Multimodal LLM Reasoning via Mitigating Visual Hallucination Self-Consistency Improves Chain of Thought Reasoning in Language Models
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 27619d99-79b7-4f59-b1fa-1f32b372487b · outbound
Thinking Before Looking: Improving Multimodal LLM Reasoning via Mitigating Visual Hallucination Chain-of-thought prompting elicits reasoning in large lan- guage models
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 96814179-5438-4fe0-9f7a-c16c0cc22857 · outbound
Thinking Before Looking: Improving Multimodal LLM Reasoning via Mitigating Visual Hallucination Self-evaluation guided beam search for reasoning
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 964677b3-b86d-47f7-97e1-7731716164d4 · outbound
Thinking Before Looking: Improving Multimodal LLM Reasoning via Mitigating Visual Hallucination Tree of thoughts: Deliberate problem solving with large language models
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7a2c6989-198b-43ee-987b-7b279199e7ff · outbound
Thinking Before Looking: Improving Multimodal LLM Reasoning via Mitigating Visual Hallucination Woodpecker: Hallucination Correction for Multimodal Large Language Models
Reference 41
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 593b5a7f-0b12-4f86-a585-7fab7a251aef · outbound
Thinking Before Looking: Improving Multimodal LLM Reasoning via Mitigating Visual Hallucination Contextual object detection with multi- modal large language models
Reference 42
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation a76c9d70-7b37-41ec-b452-3fbc1b93f645 · outbound
Thinking Before Looking: Improving Multimodal LLM Reasoning via Mitigating Visual Hallucination CoCoT: Contrastive Chain-of-Thought Prompting for Large Multimodal Models with Multiple Image Inputs
Reference 43
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0d98664b-cc5f-408c-aca6-3128f8ff547f · outbound
Thinking Before Looking: Improving Multimodal LLM Reasoning via Mitigating Visual Hallucination LLaMA-Adapter: Efficient Fine-tuning of Language Models with Zero-init Attention
Reference 44
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9f52fecb-9ad2-48ec-b77b-d136b0a4f160 · outbound
Thinking Before Looking: Improving Multimodal LLM Reasoning via Mitigating Visual Hallucination Automatic Chain of Thought Prompting in Large Language Models
Reference 45
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2f2804e1-aa08-4b14-80e7-368bba383291 · outbound
Thinking Before Looking: Improving Multimodal LLM Reasoning via Mitigating Visual Hallucination Multimodal Chain-of-Thought Reasoning in Language Models
Reference 46
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a3f11512-b227-4cc5-875e-26cbd892782d · outbound
Thinking Before Looking: Improving Multimodal LLM Reasoning via Mitigating Visual Hallucination Ddcot: Duty-distinct chain-of-thought prompting for multimodal reasoning in language models.Advances in Neu- ral Information Processing Systems, 36:5168–5191, 2023
Reference 47
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 4de40864-9826-49af-bbc7-6f277bab5103 · outbound
Thinking Before Looking: Improving Multimodal LLM Reasoning via Mitigating Visual Hallucination Image-of-Thought Prompting for Visual Reasoning Refinement in Multimodal Large Language Models
Reference 48
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation db9bf6ac-8065-4ad1-8cda-dc6bdc911ad9 · outbound
Thinking Before Looking: Improving Multimodal LLM Reasoning via Mitigating Visual Hallucination Unresolved cited work
Reference 2024
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 50ea1800-5255-440e-951f-995cfe6c9911 · inbound
Hallucination of Multimodal Large Language Models: A Survey Thinking Before Looking: Improving Multimodal LLM Reasoning via Mitigating Visual Hallucination
Reference 220
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 3b478f3d-7e97-4776-97be-1a2b1dc81a7f · inbound
Multimodal Chain-of-Thought Reasoning: A Comprehensive Survey Thinking Before Looking: Improving Multimodal LLM Reasoning via Mitigating Visual Hallucination
Reference 89
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 7f35f66e-4cbe-463e-a511-e8f97f085ad1 · inbound
MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM Thinking Before Looking: Improving Multimodal LLM Reasoning via Mitigating Visual Hallucination
Reference 85
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1235f923-a65d-4a8e-b541-feeaed84209a · inbound
Empowering Multimodal LLMs with External Tools: A Comprehensive Survey Thinking Before Looking: Improving Multimodal LLM Reasoning via Mitigating Visual Hallucination
Reference 177
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0e2aae97-76fb-4127-a7c1-4530bfc3e7ae · inbound
MEENA (PersianMMMU): Multimodal-Multilingual Educational Exams for N-level Assessment Thinking Before Looking: Improving Multimodal LLM Reasoning via Mitigating Visual Hallucination
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 140f773a-c2e3-468a-b8c4-9dc3f4a975b4 · inbound
A Comprehensive Survey on Trustworthiness in Reasoning with Large Language Models Thinking Before Looking: Improving Multimodal LLM Reasoning via Mitigating Visual Hallucination
Reference 47
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 74692ec0-193e-4668-9bc6-eeab11ca3654 · inbound
On Semiotic-Grounded Interpretive Evaluation of Generative Art Thinking Before Looking: Improving Multimodal LLM Reasoning via Mitigating Visual Hallucination
Reference 91
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 86f32005-f045-4120-a6e7-e2ddc4732861 · inbound
Learn to Think: Improving Multimodal Reasoning through Vision-Aware Self-Improvement Training Thinking Before Looking: Improving Multimodal LLM Reasoning via Mitigating Visual Hallucination
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation a612626b-f8e3-49ce-ba22-58413842c0b3 · inbound
CL-Anomaly: Layer-Adaptive Mixture-of-Experts with Multimodal Large Language Model for Continual Learning in Anomaly Detection Thinking Before Looking: Improving Multimodal LLM Reasoning via Mitigating Visual Hallucination
Reference 65
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f4cdebd0-27c0-4833-8107-4148bfc07050 · inbound
VADER: Adaptive Debiasing for Hallucination Mitigation in Video Large Language Models Thinking Before Looking: Improving Multimodal LLM Reasoning via Mitigating Visual Hallucination
Reference 99
Source-reported events for the cited work
Unavailable: canonical work link unavailable.