Pith. sign in

Paper Citation Record · LEDGER

Same Attention, Different Truths: Put Logit-Lens over Visual Attention to Detect and Mitigate LVLM Object Hallucination

As of 11 August 2026, this Paper Citation Record lists 47 of 47 outbound references and 0 inbound Pith citation observations for arXiv:2608.07302.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2608.07302 v1

Coverage vector

measured 47 of 47 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-10T10:47:44.075469Z

measured 47 of 47 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

47 of 47 outbound references displayed

  • verified exact0
  • verified fuzzy17
  • unresolved30
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 0e320b62-8865-4ece-abad-63999b4708fd · outbound

This paper cites Flamingo: a visual language model for few-shot learning.Advances in neural information processing systems, 35:23716–23736,.

Same Attention, Different Truths: Put Logit-Lens over Visual Attention to Detect and Mitigate LVLM Object Hallucination Flamingo: a visual language model for few-shot learning.Advances in neural information processing systems, 35:23716–23736,

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-10T10:47:43.891133Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T10:47:43.891133Z digest=sha256:2a82287fcdedcd6f6b6e9f01fa89f9387d7de9aa134cca22a961356bca048bf8

Observation 99d03d34-605b-4c96-9e6a-54551257cd88 · outbound

This paper cites Mitigating object hallucinations in large vision- language models with assembly of global and local attention.

Same Attention, Different Truths: Put Logit-Lens over Visual Attention to Detect and Mitigate LVLM Object Hallucination Mitigating object hallucinations in large vision- language models with assembly of global and local attention

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T10:47:44.841598Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T10:47:43.895940Z digest=sha256:97807bb15fdfd22b9c12b2cc3479c8b08d0264f60b8b746ca82c379551a82b46

Observation f8ba24a2-c3c4-4151-a77a-6592f564fb22 · outbound

This paper cites Qwen-VL: A Versatile Vision-Language Model for Understanding, Localization, Text Reading, and Beyond.

Same Attention, Different Truths: Put Logit-Lens over Visual Attention to Detect and Mitigate LVLM Object Hallucination Qwen-VL: A Versatile Vision-Language Model for Understanding, Localization, Text Reading, and Beyond

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-10T10:47:43.900405Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T10:47:43.900405Z digest=sha256:626bda6d91753d2a627e8d024eaeb611a0da861d4629cbf05fb2ddc24b15a447

Observation 82a7088a-e450-4fce-9278-68e97ad3d517 · outbound

This paper cites Hallucination of Multimodal Large Language Models: A Survey.

Same Attention, Different Truths: Put Logit-Lens over Visual Attention to Detect and Mitigate LVLM Object Hallucination Hallucination of Multimodal Large Language Models: A Survey

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-10T10:47:43.905091Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T10:47:43.905091Z digest=sha256:2a743544db46540a6d62a9463fab04274e2721e8222802aa60a2b290a2b51409

Observation 3cf3aa9d-e1d5-484b-9485-2e3fb0831961 · outbound

This paper cites Ict: Image-object cross-level trusted intervention for mitigating object halluci- nation in large vision-language models.

Same Attention, Different Truths: Put Logit-Lens over Visual Attention to Detect and Mitigate LVLM Object Hallucination Ict: Image-object cross-level trusted intervention for mitigating object halluci- nation in large vision-language models

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T10:47:44.830009Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T10:47:43.910328Z digest=sha256:d971b636ccb17f8b8167dd2557d2cca2c61c734c89997818fb0dd0ae49814286

Observation c151d0e4-ad30-4c36-9fc8-f3cbc4665155 · outbound

This paper cites Shikra: Unleashing Multimodal LLM's Referential Dialogue Magic.

Same Attention, Different Truths: Put Logit-Lens over Visual Attention to Detect and Mitigate LVLM Object Hallucination Shikra: Unleashing Multimodal LLM's Referential Dialogue Magic

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-10T10:47:43.914584Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T10:47:43.914584Z digest=sha256:a25d3add21ea3fb312fab1271cd174e038961534c2cd664ea289efdd7a180bf5

Observation 7000a756-3aa0-4606-9338-7fc8a9f5274b · outbound

This paper cites Mitigating Hallucination in Visual Language Models with Visual Supervision.

Same Attention, Different Truths: Put Logit-Lens over Visual Attention to Detect and Mitigate LVLM Object Hallucination Mitigating Hallucination in Visual Language Models with Visual Supervision

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-10T10:47:43.918681Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T10:47:43.918681Z digest=sha256:ea2f2abd2ef6ccfb82569f582aac826c2533818464c70dd0f37540c9849c5ffa

Observation b11a80a3-aaa8-412a-8a3a-176479ab9f85 · outbound

This paper cites Internvl: Scaling up vision foundation mod- els and aligning for generic visual-linguistic tasks.

Same Attention, Different Truths: Put Logit-Lens over Visual Attention to Detect and Mitigate LVLM Object Hallucination Internvl: Scaling up vision foundation mod- els and aligning for generic visual-linguistic tasks

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-10T10:47:43.922861Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T10:47:43.922861Z digest=sha256:62f234ca0afbf062f9bb1c81eaf90b375ade890ca751917f283ddf203f14330d

Observation e89ab366-f85c-431a-8c3a-7467c7ac55c2 · outbound

This paper cites HALC: Object Hallucination Reduction via Adaptive Focal-Contrast Decoding.

Same Attention, Different Truths: Put Logit-Lens over Visual Attention to Detect and Mitigate LVLM Object Hallucination HALC: Object Hallucination Reduction via Adaptive Focal-Contrast Decoding

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-10T10:47:43.926644Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T10:47:43.926644Z digest=sha256:5e676042d481055b66c63e9b7c26ecc2e3dced942201369a661892b6bb83db71

Observation d4dde930-0bf6-42af-87fb-bbe478075ca4 · outbound

This paper cites Instructblip: Towards general-purpose vision- language models with instruction tuning.Advances in neural information processing systems, 36:49250–49267, 2023.

Same Attention, Different Truths: Put Logit-Lens over Visual Attention to Detect and Mitigate LVLM Object Hallucination Instructblip: Towards general-purpose vision- language models with instruction tuning.Advances in neural information processing systems, 36:49250–49267, 2023

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-10T10:47:43.930595Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T10:47:43.930595Z digest=sha256:17227e3e55aade4713f2897454f784f0aa6566e94d920993dba2c4cc9ca41ff0

Observation 60cb3f27-032c-4303-8ea5-e7a8d8d689d4 · outbound

This paper cites Damro: Dive into the attention mechanism of lvlm to re- duce object hallucination.arXiv preprint arXiv:2410.04514,.

Same Attention, Different Truths: Put Logit-Lens over Visual Attention to Detect and Mitigate LVLM Object Hallucination Damro: Dive into the attention mechanism of lvlm to re- duce object hallucination.arXiv preprint arXiv:2410.04514,

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-10T10:47:43.934722Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T10:47:43.934722Z digest=sha256:58e8e87f7eb5595637be99d29ea071ab21b038260012c509d90911a60aad312d

Observation 86a2d1d9-cf6e-4cc0-962f-6032434cbb14 · outbound

This paper cites Opera: Alleviating hallucination in multi- modal large language models via over-trust penalty and retrospection-allocation.

Same Attention, Different Truths: Put Logit-Lens over Visual Attention to Detect and Mitigate LVLM Object Hallucination Opera: Alleviating hallucination in multi- modal large language models via over-trust penalty and retrospection-allocation

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T10:47:44.803823Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T10:47:43.938593Z digest=sha256:fbe520d5b78e55cf35c187493ad10bc5ebb7c570f2a0b8a106efaaa126a347df

Observation 00a1ab87-7705-4d79-bfc7-4bf232784739 · outbound

This paper cites Hallucination augmented contrastive learn- ing for multimodal large language model.

Same Attention, Different Truths: Put Logit-Lens over Visual Attention to Detect and Mitigate LVLM Object Hallucination Hallucination augmented contrastive learn- ing for multimodal large language model

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T10:47:44.791239Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T10:47:43.942300Z digest=sha256:0da75bd680861783bbf69372a50875a76854296dddb6937e0bddc46429d0e737

Observation 615a9787-1c90-4108-8fb9-c20c8e3c0411 · outbound

This paper cites Interpreting and Editing Vision-Language Representations to Mitigate Hallucinations.

Same Attention, Different Truths: Put Logit-Lens over Visual Attention to Detect and Mitigate LVLM Object Hallucination Interpreting and Editing Vision-Language Representations to Mitigate Hallucinations

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-10T10:47:43.945970Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T10:47:43.945970Z digest=sha256:980cd06e9e76891dbfbdc800c4f048c3cfe48200c90a274bfb685fd0b6c0a6ee

Observation fcda5f3b-5dad-45a7-b2c4-e0e8ca6eed1a · outbound

This paper cites Devils in middle layers of large vision- language models: Interpreting, detecting and mitigating ob- ject hallucinations via attention lens.

Same Attention, Different Truths: Put Logit-Lens over Visual Attention to Detect and Mitigate LVLM Object Hallucination Devils in middle layers of large vision- language models: Interpreting, detecting and mitigating ob- ject hallucinations via attention lens

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T10:47:44.780119Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T10:47:43.949886Z digest=sha256:7b3a0cfb46b16234388d8b2094fc4a4c25bfce0ccecda4bbbfdecf91cbce0a39

Observation 2dfebab3-5a67-4cf4-8e2f-57ef29db632b · outbound

This paper cites What’s in the im- age? a deep-dive into the vision of vision language models.

Same Attention, Different Truths: Put Logit-Lens over Visual Attention to Detect and Mitigate LVLM Object Hallucination What’s in the im- age? a deep-dive into the vision of vision language models

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T10:47:44.768877Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T10:47:43.953525Z digest=sha256:5a8853bc7a04e71dd241761bef96d5435614f2ac1b3157284c8e543cdc665eee

Observation bc3af201-0a87-4966-a990-2f6c2f8c9e5d · outbound

This paper cites See What You Are Told: Visual Attention Sink in Large Multimodal Models.

Same Attention, Different Truths: Put Logit-Lens over Visual Attention to Detect and Mitigate LVLM Object Hallucination See What You Are Told: Visual Attention Sink in Large Multimodal Models

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-10T10:47:43.957166Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T10:47:43.957166Z digest=sha256:ab6484ef27d5843a59210a88869a338969e2564a0de9b48a2386cef5a6661abb

Observation 2deb19e1-8498-4016-99bc-2ff7fe9c96b5 · outbound

This paper cites Mitigating object hal- lucinations in large vision-language models through visual contrastive decoding.

Same Attention, Different Truths: Put Logit-Lens over Visual Attention to Detect and Mitigate LVLM Object Hallucination Mitigating object hal- lucinations in large vision-language models through visual contrastive decoding

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T10:47:44.756687Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T10:47:43.960888Z digest=sha256:b3fd0dfc1968c012f4c255304d66471bdda93fef6fbf5fbae7e1ac50dd67cdd3

Observation dee3fdc5-0835-45d7-971e-9ba4096b2e8a · outbound

This paper cites Blip-2: Bootstrapping language-image pre-training with frozen image encoders and large language models.

Same Attention, Different Truths: Put Logit-Lens over Visual Attention to Detect and Mitigate LVLM Object Hallucination Blip-2: Bootstrapping language-image pre-training with frozen image encoders and large language models

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-10T10:47:43.964835Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T10:47:43.964835Z digest=sha256:db69116e0345d55ccf084dc6d00c55443e03383068d3edab46c3b8b42fab79c6

Observation 5e31a5d3-4f45-4260-a211-e3fe9936d7b9 · outbound

This paper cites Mvbench: A comprehensive multi-modal video understand- ing benchmark.

Same Attention, Different Truths: Put Logit-Lens over Visual Attention to Detect and Mitigate LVLM Object Hallucination Mvbench: A comprehensive multi-modal video understand- ing benchmark

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-10T10:47:43.968690Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T10:47:43.968690Z digest=sha256:fbb96f720ba63be48b3d74993d394bb36ec7384af415d930e1fe80f291eb8163

Observation c87cbf62-5bc7-4ca1-a1b5-6d09e73ae0b9 · outbound

This paper cites Contrastive decoding: Open-ended text genera- tion as optimization.

Same Attention, Different Truths: Put Logit-Lens over Visual Attention to Detect and Mitigate LVLM Object Hallucination Contrastive decoding: Open-ended text genera- tion as optimization

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T10:47:44.731096Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T10:47:43.972351Z digest=sha256:83c78b0c3d4752824c7fc8ea6b3ab6a1e11f48b1a3ac57f5da0d77c5bf2947a9

Observation dcc333b6-e883-47d0-ab57-4aae01192727 · outbound

This paper cites Evaluating Object Hallucination in Large Vision-Language Models.

Same Attention, Different Truths: Put Logit-Lens over Visual Attention to Detect and Mitigate LVLM Object Hallucination Evaluating Object Hallucination in Large Vision-Language Models

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-10T10:47:43.975914Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T10:47:43.975914Z digest=sha256:96915d2df3512526187d4d79c46d84f1be3eab9639a3ae56da88961a4ef6cadb

Observation eabe12cc-a553-4028-b6fd-31e4c2dd20f1 · outbound

This paper cites Microsoft coco: Common objects in context.

Same Attention, Different Truths: Put Logit-Lens over Visual Attention to Detect and Mitigate LVLM Object Hallucination Microsoft coco: Common objects in context

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-10T10:47:43.979840Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T10:47:43.979840Z digest=sha256:10456685f9da70bdfedf7a29af07f733f1613b9a6a1b0e4ccdca87d8dc9c0896

Observation 887a82ea-f4bb-42a5-adde-1e19b3be83ff · outbound

This paper cites Mitigating Hallucination in Large Multi-Modal Models via Robust Instruction Tuning.

Same Attention, Different Truths: Put Logit-Lens over Visual Attention to Detect and Mitigate LVLM Object Hallucination Mitigating Hallucination in Large Multi-Modal Models via Robust Instruction Tuning

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-10T10:47:43.983427Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T10:47:43.983427Z digest=sha256:1ad7d16e50ee6289054b3609dd1147a7486b6fe89268084e818d21b6468f3aab

Observation fa1b95be-cf8f-4b06-bee2-bf17ed85f7a0 · outbound

This paper cites Visual instruction tuning.Advances in neural information processing systems, 36:34892–34916, 2023.

Same Attention, Different Truths: Put Logit-Lens over Visual Attention to Detect and Mitigate LVLM Object Hallucination Visual instruction tuning.Advances in neural information processing systems, 36:34892–34916, 2023

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-10T10:47:43.987399Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T10:47:43.987399Z digest=sha256:ef85a1fd35d1804a8792245f667e71eb0dd82ba5a6aa645f73e5a30ad9e161e9

Observation ab154c61-d930-40b6-8d2b-c4691ab32fa9 · outbound

This paper cites Improved baselines with visual instruction tuning.

Same Attention, Different Truths: Put Logit-Lens over Visual Attention to Detect and Mitigate LVLM Object Hallucination Improved baselines with visual instruction tuning

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T10:47:44.702842Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T10:47:43.990877Z digest=sha256:a18072fcc1611cd1d3fd2acf132a7672d2ca7bfbc5f3470ee4560154397db176

Observation e3ca8da4-6f46-4fb1-b164-22c40ec6631c · outbound

This paper cites A Survey on Hallucination in Large Vision-Language Models.

Same Attention, Different Truths: Put Logit-Lens over Visual Attention to Detect and Mitigate LVLM Object Hallucination A Survey on Hallucination in Large Vision-Language Models

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-10T10:47:43.994864Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T10:47:43.994864Z digest=sha256:577f33446ab133d04d141167b12da4482ab59e903e6c9e3933443a0361cb16e0

Observation 53c0377c-6c8c-43b3-9eef-9b2059253bd3 · outbound

This paper cites Paying more at- tention to image: A training-free method for alleviating hal- lucination in lvlms.

Same Attention, Different Truths: Put Logit-Lens over Visual Attention to Detect and Mitigate LVLM Object Hallucination Paying more at- tention to image: A training-free method for alleviating hal- lucination in lvlms

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T10:47:44.689548Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T10:47:43.998702Z digest=sha256:f9b9eccf32645a8354dbdeeb938ca874915a26da3df0d9f992566b324b37c08c

Observation 88eb18a9-d520-4dd6-9d9f-56240c02f130 · outbound

This paper cites Alleviating hallucinations in large vision- language models through hallucination-induced optimiza- tion.Advances in Neural Information Processing Systems, 37:122811–122832, 2024.

Same Attention, Different Truths: Put Logit-Lens over Visual Attention to Detect and Mitigate LVLM Object Hallucination Alleviating hallucinations in large vision- language models through hallucination-induced optimiza- tion.Advances in Neural Information Processing Systems, 37:122811–122832, 2024

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T10:47:44.676793Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T10:47:44.002374Z digest=sha256:f7d6690f141924a4592716f21a43974fc62f92811280031034fcdb5b58d23abf

Observation 9f54d766-b08b-440c-9f24-7a90343b4901 · outbound

This paper cites Wordnet: a lexical database for english.

Same Attention, Different Truths: Put Logit-Lens over Visual Attention to Detect and Mitigate LVLM Object Hallucination Wordnet: a lexical database for english

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T10:47:44.664276Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T10:47:44.006316Z digest=sha256:ab37d080a13e88f21b4f4a001828110b5f4d0e8a3c20e20a3f7949c8a8667f64

Observation c721d87f-671f-40b3-b4d1-e4b82b481a6d · outbound

This paper cites Interpreting gpt: The logit lens.

Same Attention, Different Truths: Put Logit-Lens over Visual Attention to Detect and Mitigate LVLM Object Hallucination Interpreting gpt: The logit lens

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T10:47:44.653301Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T10:47:44.010190Z digest=sha256:1b0f8d4f52d90416b52d8d51129709db8ec7d1f094fd7ff657f6c9133cba3391

Observation 8790dd97-557f-4dd7-b86d-cdfdb9a45777 · outbound

This paper cites Learning transferable visual models from natural language supervi- sion.

Same Attention, Different Truths: Put Logit-Lens over Visual Attention to Detect and Mitigate LVLM Object Hallucination Learning transferable visual models from natural language supervi- sion

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-10T10:47:44.014328Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T10:47:44.014328Z digest=sha256:4320b763a2752c70a610c4751ffc21ef61ff9ac76cc20c473e325a6fb587d1c4

Observation f4768a1b-8578-41a1-89c9-2e7b3fa691ed · outbound

This paper cites Object Hallucination in Image Captioning.

Same Attention, Different Truths: Put Logit-Lens over Visual Attention to Detect and Mitigate LVLM Object Hallucination Object Hallucination in Image Captioning

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-10T10:47:44.018119Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T10:47:44.018119Z digest=sha256:6f4df99a6d64cb41fcd12b2bcc4cae92605eaa9d8f9c1e4a2ccbcd6b4b8d4eae

Observation 44ff7873-12ee-4b4f-a979-a48396d2dd0a · outbound

This paper cites Mitigating Object Hallucination in MLLMs via Data-augmented Phrase-level Alignment.

Same Attention, Different Truths: Put Logit-Lens over Visual Attention to Detect and Mitigate LVLM Object Hallucination Mitigating Object Hallucination in MLLMs via Data-augmented Phrase-level Alignment

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-10T10:47:44.022178Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T10:47:44.022178Z digest=sha256:efeee885b2570e8a907bf2f146603e2ec22fa5bbc7e9ed4efc2a1ba2d0c67381

Observation db768d3b-9524-429e-b314-8d5741a8c656 · outbound

This paper cites Overconfidence in LLM-as-a-Judge: Diagnosis and Confidence-Driven Solution.

Same Attention, Different Truths: Put Logit-Lens over Visual Attention to Detect and Mitigate LVLM Object Hallucination Overconfidence in LLM-as-a-Judge: Diagnosis and Confidence-Driven Solution

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-10T10:47:44.026423Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T10:47:44.026423Z digest=sha256:6979c07501f0bdc01bdd745551c6465dc21d72c7a03d033250fc21afbc482833

Observation 6162e675-7f02-4ab3-8c82-3394dccfaeac · outbound

This paper cites Eyes wide shut? exploring the visual shortcomings of multimodal llms.

Same Attention, Different Truths: Put Logit-Lens over Visual Attention to Detect and Mitigate LVLM Object Hallucination Eyes wide shut? exploring the visual shortcomings of multimodal llms

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T10:47:44.634588Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T10:47:44.030588Z digest=sha256:bc9ae50fda22c0b54e0ed77c1b22dcaa797df77ea3afe9c7d47639dd71499268

Observation 649df8aa-6029-4b67-a3ca-17bab84bd5be · outbound

This paper cites MLLM can see? Dynamic Correction Decoding for Hallucination Mitigation.

Same Attention, Different Truths: Put Logit-Lens over Visual Attention to Detect and Mitigate LVLM Object Hallucination MLLM can see? Dynamic Correction Decoding for Hallucination Mitigation

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-10T10:47:44.034645Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T10:47:44.034645Z digest=sha256:36768400a7e580f06946aa808d487c817f16f3370feef8ef46f0efac0fe41290

Observation 55a95a22-9657-428c-8997-9ae74d884b58 · outbound

This paper cites AMBER: An LLM-free Multi-dimensional Benchmark for MLLMs Hallucination Evaluation.

Same Attention, Different Truths: Put Logit-Lens over Visual Attention to Detect and Mitigate LVLM Object Hallucination AMBER: An LLM-free Multi-dimensional Benchmark for MLLMs Hallucination Evaluation

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-10T10:47:44.038832Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T10:47:44.038832Z digest=sha256:9717742843e126e96a223bd8f16d3b305a230b24cb923558e5cd7108ebd75c35

Observation 910c0aa5-a815-46c1-83b3-d77486dd8de7 · outbound

This paper cites Qwen2-VL: Enhancing Vision-Language Model's Perception of the World at Any Resolution.

Same Attention, Different Truths: Put Logit-Lens over Visual Attention to Detect and Mitigate LVLM Object Hallucination Qwen2-VL: Enhancing Vision-Language Model's Perception of the World at Any Resolution

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-10T10:47:44.043045Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T10:47:44.043045Z digest=sha256:06c96c3c4adab4153db5b830167d3a09848e323815c7d4725d8af29fa4bd163a

Observation fe3929af-e1b9-4de8-a0b4-6ce2f51ee134 · outbound

This paper cites When Language Overrules: Revealing Text Dominance in Multimodal Large Language Models.

Same Attention, Different Truths: Put Logit-Lens over Visual Attention to Detect and Mitigate LVLM Object Hallucination When Language Overrules: Revealing Text Dominance in Multimodal Large Language Models

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-10T10:47:44.047025Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T10:47:44.047025Z digest=sha256:31e3db7d1bc8fc44a821559e4154e2d76343e893d8e08ab2cf82a83450bb6e19

Observation 4d20456f-62e9-461e-998d-1b1f65351d6d · outbound

This paper cites Mitigating hallucinations in large vision- language models via dpo: On-policy data hold the key.

Same Attention, Different Truths: Put Logit-Lens over Visual Attention to Detect and Mitigate LVLM Object Hallucination Mitigating hallucinations in large vision- language models via dpo: On-policy data hold the key

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T10:47:44.622766Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T10:47:44.051227Z digest=sha256:2d28230a384d2d8f05a218171213b57d7e02ececfa68c52595dfd2495a7d32df

Observation 685cf51b-407d-4d66-8d5b-e202331a509f · outbound

This paper cites mplug- owl2: Revolutionizing multi-modal large language model with modality collaboration.

Same Attention, Different Truths: Put Logit-Lens over Visual Attention to Detect and Mitigate LVLM Object Hallucination mplug- owl2: Revolutionizing multi-modal large language model with modality collaboration

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-10T10:47:44.055257Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T10:47:44.055257Z digest=sha256:89511c920bee88b21b0dedea6940254a1ca8bf992938510db20266364dd91072

Observation d4c5925a-c39d-41ea-8266-012b2ff406ff · outbound

This paper cites Clearsight: Vi- sual signal enhancement for object hallucination mitigation in multimodal large language models.

Same Attention, Different Truths: Put Logit-Lens over Visual Attention to Detect and Mitigate LVLM Object Hallucination Clearsight: Vi- sual signal enhancement for object hallucination mitigation in multimodal large language models

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-10T10:47:44.059547Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T10:47:44.059547Z digest=sha256:8261d981bb3c7e5879bd2e584e4f091009f79023a00cca1da858c812446d6fb1

Observation 02458c0e-89e6-48e6-89a6-b0df3c721b8c · outbound

This paper cites Rlhf-v: Towards trustworthy mllms via behavior alignment from fine-grained correctional hu- man feedback.

Same Attention, Different Truths: Put Logit-Lens over Visual Attention to Detect and Mitigate LVLM Object Hallucination Rlhf-v: Towards trustworthy mllms via behavior alignment from fine-grained correctional hu- man feedback

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T10:47:44.593865Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T10:47:44.063506Z digest=sha256:4b5fac023c05ab89145f9771655f6591bffc2cd3e0301f4c25fbb8daa2cb709d

Observation 39a2c8f4-0a3a-4fac-9053-2ff1554b32bc · outbound

This paper cites Mitigating object hallucination in large vision-language models via classifier-free guidance.arXiv e-prints, pages arXiv–2402, 2024.

Same Attention, Different Truths: Put Logit-Lens over Visual Attention to Detect and Mitigate LVLM Object Hallucination Mitigating object hallucination in large vision-language models via classifier-free guidance.arXiv e-prints, pages arXiv–2402, 2024

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T10:47:44.580796Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T10:47:44.067679Z digest=sha256:0b48d07e758c08c9f481e82d823c1e70ed53ecce969c4a03016880a41c2773fa

Observation b4d8f383-ae8a-4296-ae4a-e8c09fd8a0e8 · outbound

This paper cites Beyond Hallucinations: Enhancing LVLMs through Hallucination-Aware Direct Preference Optimization.

Same Attention, Different Truths: Put Logit-Lens over Visual Attention to Detect and Mitigate LVLM Object Hallucination Beyond Hallucinations: Enhancing LVLMs through Hallucination-Aware Direct Preference Optimization

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-10T10:47:44.071502Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T10:47:44.071502Z digest=sha256:d0ad47711144ed98de81350f261958498efb01157ec5e4bd8a0b55ad29df99a8

Observation 01dad650-bf56-4aed-9657-a284790b120b · outbound

This paper cites Analyzing and Mitigating Object Hallucination in Large Vision-Language Models.

Same Attention, Different Truths: Put Logit-Lens over Visual Attention to Detect and Mitigate LVLM Object Hallucination Analyzing and Mitigating Object Hallucination in Large Vision-Language Models

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-10T10:47:44.075469Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T10:47:44.075469Z digest=sha256:c842ace5fcc0d7d0cd6f15fd5b8a1ef2daaf68c9f0cf4dfee75177998058864b

Pith citing papers

No inbound Pith citation observations are available.