Pith. sign in

Paper Citation Record · LEDGER

Seeing It or Not? Interpretable Vision-aware Latent Steering to Mitigate Object Hallucinations

As of 23 August 2026, this Paper Citation Record lists 64 of 64 outbound references and 0 inbound Pith citation observations for arXiv:2505.17812.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.17812 v1

Coverage vector

measured 64 of 64 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T14:45:32.061891Z

measured 64 of 64 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-23T06:30:58.430688+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

64 of 64 outbound references displayed

  • verified exact3
  • verified fuzzy31
  • unresolved29
  • parse uncertain0
  • malformed identifier1
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 882aa6cd-9854-4dfc-8100-a9006e0ca4f8 · outbound

This paper cites LLaMA: Open and Efficient Foundation Language Models.

Seeing It or Not? Interpretable Vision-aware Latent Steering to Mitigate Object Hallucinations LLaMA: Open and Efficient Foundation Language Models

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-07T14:45:26.861270Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:45:26.861270Z digest=sha256:fc86ce29313f72acef5c07727362cece679ce086e17cff4aef013cc024a51c27

Observation 737a10df-2d22-40cb-bfce-2668d47766e6 · outbound

This paper cites Llama 2: Open Foundation and Fine-Tuned Chat Models.

Seeing It or Not? Interpretable Vision-aware Latent Steering to Mitigate Object Hallucinations Llama 2: Open Foundation and Fine-Tuned Chat Models

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-07T14:45:26.929012Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:45:26.929012Z digest=sha256:c698a8fa332cd46b2288f5c53145beb0b352adb5c9ca5f04598594eea9fff719

Observation 81c64c90-6c7e-4622-bfbb-60dfd6c91d37 · outbound

This paper cites Visual instruction tuning.Adv.

Seeing It or Not? Interpretable Vision-aware Latent Steering to Mitigate Object Hallucinations Visual instruction tuning.Adv

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:45:39.249729Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-07T14:45:27.001968Z digest=sha256:ec753c796ea0198fc632d3b2f7e364af9d7425cffc805f3e9f70364afc0b79ee

Observation 4455e4e8-22fe-4da9-847f-c24901bf3ee4 · outbound

This paper cites Improved Baselines with Visual Instruction Tuning.

Seeing It or Not? Interpretable Vision-aware Latent Steering to Mitigate Object Hallucinations Improved Baselines with Visual Instruction Tuning

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-07T14:45:27.094656Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:45:27.094656Z digest=sha256:139d4b09366ce3019aaa7a9ac09da169f88aa98a3d823146c2c710ed20e3b650

Observation c8a21e09-834d-476c-9bdd-064d5f45d427 · outbound

This paper cites InstructBLIP: Towards General-purpose Vision-Language Models with Instruction Tuning.

Seeing It or Not? Interpretable Vision-aware Latent Steering to Mitigate Object Hallucinations InstructBLIP: Towards General-purpose Vision-Language Models with Instruction Tuning

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-07T14:45:27.170843Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:45:27.170843Z digest=sha256:c841cbbe6ad85c71f9da23bae5c602f6161829ebab691e71b0d0a0d876cc4974

Observation 8f887f16-e771-46b9-8839-be6850db0e26 · outbound

This paper cites MiniGPT-4: Enhancing Vision-Language Understanding with Advanced Large Language Models.

Seeing It or Not? Interpretable Vision-aware Latent Steering to Mitigate Object Hallucinations MiniGPT-4: Enhancing Vision-Language Understanding with Advanced Large Language Models

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-07T14:45:27.251904Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:45:27.251904Z digest=sha256:e8c7e9002fcfcea3fca49dbaa4b22c643048f27ecf894071a59104cbb56ebee7

Observation 9e209bd6-89d7-4769-830f-56caefcb9c6f · outbound

This paper cites Qwen Technical Report.

Seeing It or Not? Interpretable Vision-aware Latent Steering to Mitigate Object Hallucinations Qwen Technical Report

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-07T14:45:27.325359Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:45:27.325359Z digest=sha256:adba8a68c6a98ded2d49d8033e4836ba8bec9282bdb9b16294bf40fd4be5ff0c

Observation cfe75358-ff73-4c7d-8cbe-b2cfef5b7b6e · outbound

This paper cites Qwen2-VL: Enhancing Vision-Language Model's Perception of the World at Any Resolution.

Seeing It or Not? Interpretable Vision-aware Latent Steering to Mitigate Object Hallucinations Qwen2-VL: Enhancing Vision-Language Model's Perception of the World at Any Resolution

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-07T14:45:27.396913Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:45:27.396913Z digest=sha256:5bb265d870fd1637f2280a83c1da0fb9183d5bb028f1f1716042393af8020f8e

Observation f1d4b41e-4c28-4897-82a9-ea75f1ccdfe5 · outbound

This paper cites Hallucination of Multimodal Large Language Models: A Survey.

Seeing It or Not? Interpretable Vision-aware Latent Steering to Mitigate Object Hallucinations Hallucination of Multimodal Large Language Models: A Survey

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-07T14:45:27.497184Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:45:27.497184Z digest=sha256:977834da815dba06c3ca4fb2e0d7739cfdf935898d36dbe7d9ea99e9f2f8371b

Observation 66e2ef9a-7aaf-41da-9155-14405f3ba299 · outbound

This paper cites Nullu: Mitigating object hallucinations in large vision-language models via halluspace projection.

Seeing It or Not? Interpretable Vision-aware Latent Steering to Mitigate Object Hallucinations Nullu: Mitigating object hallucinations in large vision-language models via halluspace projection

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:45:39.061415Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-07T14:45:27.582588Z digest=sha256:5e2ef806450054b1abec33e353e1df6f9dd0eefb3ac8a17fe5a6d4c00adfd6b7

Observation f12fe377-086e-4152-a5f1-8a0c152760c5 · outbound

This paper cites Truthprint: Mitigating lvlm object hallucination via latent truthful-guided pre-intervention.arXiv preprint arXiv:2503.10602, 2025.

Seeing It or Not? Interpretable Vision-aware Latent Steering to Mitigate Object Hallucinations Truthprint: Mitigating lvlm object hallucination via latent truthful-guided pre-intervention.arXiv preprint arXiv:2503.10602, 2025

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-07T14:45:27.675017Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:45:27.675017Z digest=sha256:d0309b7e758da57ddb57c73bbfbbce6b06608d606aa252a3b213c0937e5664c4

Observation 0473397b-fec5-45da-b3c0-13bccfee3515 · outbound

This paper cites Analyzing and mitigating object hallucination in large vision-language models.

Seeing It or Not? Interpretable Vision-aware Latent Steering to Mitigate Object Hallucinations Analyzing and mitigating object hallucination in large vision-language models

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:45:38.889223Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-07T14:45:27.767172Z digest=sha256:4b5bd352f394e736a6c84afbe345c3b917c6a1bbfa5f0c883c46159bbbdd9f34

Observation 7bd6d8df-457b-42e7-abf7-ea2e01419e83 · outbound

This paper cites Mitigating Hallucination in Large Multi-Modal Models via Robust Instruction Tuning.

Seeing It or Not? Interpretable Vision-aware Latent Steering to Mitigate Object Hallucinations Mitigating Hallucination in Large Multi-Modal Models via Robust Instruction Tuning

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-07T14:45:27.838292Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:45:27.838292Z digest=sha256:9ae2c4980a6fb81b719879dec72b9ccde1c1dc11081f094d3c06c75f9973c970

Observation c1a8a98d-b379-472b-9222-2525fcf3a5c7 · outbound

This paper cites Hallucination augmented contrastive learning for multimodal large language model.

Seeing It or Not? Interpretable Vision-aware Latent Steering to Mitigate Object Hallucinations Hallucination augmented contrastive learning for multimodal large language model

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:45:38.723563Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-07T14:45:27.945183Z digest=sha256:61faff60c45f30c6f84248bd4476a2895d1a1300bdfc1b35918b5be15a3db746

Observation 4c94a0d3-e0d6-4f2f-b22b-47cf3c77c1a1 · outbound

This paper cites Exposing and mitigating spurious correlations for cross-modal retrieval.

Seeing It or Not? Interpretable Vision-aware Latent Steering to Mitigate Object Hallucinations Exposing and mitigating spurious correlations for cross-modal retrieval

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:45:38.528588Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-07T14:45:28.024239Z digest=sha256:2881d794f06b3fb0582438a9daecab09b82836fbb880a62e929c14d0841a06cf

Observation 5671c12d-a4aa-4613-a229-587492166ecc · outbound

This paper cites Mitigating object hallucinations in large vision-language models through visual contrastive decoding.

Seeing It or Not? Interpretable Vision-aware Latent Steering to Mitigate Object Hallucinations Mitigating object hallucinations in large vision-language models through visual contrastive decoding

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:45:38.339332Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-07T14:45:28.107255Z digest=sha256:17765fe24a53cf12c512131bf1f1a9805b8b0e983240aff1b44ee005fd342cf0

Observation 196a4003-abf4-40d3-8f5e-25063d5a3327 · outbound

This paper cites Debiasing large visual language models.

Seeing It or Not? Interpretable Vision-aware Latent Steering to Mitigate Object Hallucinations Debiasing large visual language models

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:45:38.147222Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-07T14:45:28.181188Z digest=sha256:a00bde8b0ad063cf9f06ee8c18928da228042507b06c66e9c77cb30773bb565c

Observation 65bf3d5c-5958-41c4-ae01-28e195126a20 · outbound

This paper cites Halc: Object hallucination reduction via adaptive focal-contrast decoding.

Seeing It or Not? Interpretable Vision-aware Latent Steering to Mitigate Object Hallucinations Halc: Object hallucination reduction via adaptive focal-contrast decoding

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:45:37.936158Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-07T14:45:28.261744Z digest=sha256:62b2258ea0f83ef3d4033b5ef7464a36900d19104a8942a442cb6f6fded6d3f4

Observation ad860591-dd20-4f99-9944-9c9f1db318f6 · outbound

This paper cites ICT: Image-Object Cross-Level Trusted Intervention for Mitigating Object Hallucination in Large Vision-Language Models.

Seeing It or Not? Interpretable Vision-aware Latent Steering to Mitigate Object Hallucinations ICT: Image-Object Cross-Level Trusted Intervention for Mitigating Object Hallucination in Large Vision-Language Models

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-07T14:45:28.344201Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:45:28.344201Z digest=sha256:bebf86fc9a55726f6a95dc88914c0e94dd5e15370c315b1226e5a5877f29cd19

Observation 25ab535f-ea73-4a22-8646-e069e69f639a · outbound

This paper cites Reducing hallucinations in large vision-language models via latent space steering.

Seeing It or Not? Interpretable Vision-aware Latent Steering to Mitigate Object Hallucinations Reducing hallucinations in large vision-language models via latent space steering

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:45:37.709424Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-07T14:45:28.404847Z digest=sha256:fd601cea8d2c7203e999141b5aaef5c82e2a49725c83da257d216abaa87e363a

Observation 0cfe0c03-bf00-4cff-95a9-3fae186b43fb · outbound

This paper cites Where do Large Vision-Language Models Look at when Answering Questions?.

Seeing It or Not? Interpretable Vision-aware Latent Steering to Mitigate Object Hallucinations Where do Large Vision-Language Models Look at when Answering Questions?

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-07T14:45:28.465442Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:45:28.465442Z digest=sha256:6f9d357698449e12048a377775577b9e2968ca7fafcf0d31fd61e3c0805f7078

Observation 49f3903b-7a7a-4ab0-9e11-feaee20efb4a · outbound

This paper cites Lvlm-intrepret: An interpretability tool for large vision-language models, 2024.

Seeing It or Not? Interpretable Vision-aware Latent Steering to Mitigate Object Hallucinations Lvlm-intrepret: An interpretability tool for large vision-language models, 2024

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:45:37.516328Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-07T14:45:28.517311Z digest=sha256:6621583511b5743a0e699c898665a232acf19d59f3b9a7c550e06fff8730c466

Observation 26f19368-69be-4297-ad00-dc9f67a84f3a · outbound

This paper cites See What You Are Told: Visual Attention Sink in Large Multimodal Models.

Seeing It or Not? Interpretable Vision-aware Latent Steering to Mitigate Object Hallucinations See What You Are Told: Visual Attention Sink in Large Multimodal Models

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-07T14:45:28.622401Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:45:28.622401Z digest=sha256:0b46e306fd7d696623182639ffda256b10d839811d03c9600ef04266f0a34beb

Observation 3f9dcb33-210b-4d5c-bbad-5ea4f33e9034 · outbound

This paper cites Vision Transformers Need Registers.

Seeing It or Not? Interpretable Vision-aware Latent Steering to Mitigate Object Hallucinations Vision Transformers Need Registers

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-07T14:45:28.704062Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:45:28.704062Z digest=sha256:e5cd85b3221a8b9ecc2a4261d41c0d76a503bb4f1fa35081dd2dcff38c8111c2

Observation bb0f94bf-7c2a-4275-9518-11fbd28bcacf · outbound

This paper cites Massive Activations in Large Language Models.

Seeing It or Not? Interpretable Vision-aware Latent Steering to Mitigate Object Hallucinations Massive Activations in Large Language Models

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-07T14:45:28.759191Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:45:28.759191Z digest=sha256:b52cfb6c0148ab6588f6fdc4da169b8dd36e6983eed2c20f835d959a1191543d

Observation ac162876-881f-4b81-8f13-701f6b5eeb37 · outbound

This paper cites Object Hallucination in Image Captioning.

Seeing It or Not? Interpretable Vision-aware Latent Steering to Mitigate Object Hallucinations Object Hallucination in Image Captioning

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-07T14:45:28.823154Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:45:28.823154Z digest=sha256:2375428858f08632358e15c44d8d552ac5765fd134f88009fcd2ea243bc374a6

Observation d6fcddd8-1821-40b3-b54b-4f0175019e62 · outbound

This paper cites mplug-owl2: Revolutionizing multi-modal large language model with modality collaboration.

Seeing It or Not? Interpretable Vision-aware Latent Steering to Mitigate Object Hallucinations mplug-owl2: Revolutionizing multi-modal large language model with modality collaboration

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:45:37.289552Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-07T14:45:28.894823Z digest=sha256:22b7b860d95cd5300233a395769ed83b544437a8a555506b1a295acab0cfe29b

Observation de3e5d8d-fea9-4b4c-892f-d99344a2f540 · outbound

This paper cites mPLUG-Owl3: Towards Long Image-Sequence Understanding in Multi-Modal Large Language Models.

Seeing It or Not? Interpretable Vision-aware Latent Steering to Mitigate Object Hallucinations mPLUG-Owl3: Towards Long Image-Sequence Understanding in Multi-Modal Large Language Models

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-07T14:45:28.975925Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:45:28.975925Z digest=sha256:ffccc1cba86a7e6738c5b3208ca71533166b81f3337858893b4e9f3ff176db39

Observation 9ab16929-c20f-4955-b6bc-ff8ad009c7d1 · outbound

This paper cites Llava-phi: Efficient multi-modal assistant with small language model.

Seeing It or Not? Interpretable Vision-aware Latent Steering to Mitigate Object Hallucinations Llava-phi: Efficient multi-modal assistant with small language model

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:45:37.107790Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-07T14:45:29.029305Z digest=sha256:d8fa304fbb4f1b0330fc7ceb0332231a48840fcb8952cbb7790364287390367b

Observation fc8a6a40-c744-42de-9ae0-c7cf38e76412 · outbound

This paper cites DeepSeek-VL: Towards Real-World Vision-Language Understanding.

Seeing It or Not? Interpretable Vision-aware Latent Steering to Mitigate Object Hallucinations DeepSeek-VL: Towards Real-World Vision-Language Understanding

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-07T14:45:29.111032Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:45:29.111032Z digest=sha256:62467a5debfcc484de96c68aa3da837a25f58c0abea50294ae61cc77422834bb

Observation cdd349ec-b545-4bf8-9229-43c1bfdc52b9 · outbound

This paper cites Detecting and preventing hallucinations in large vision language models.

Seeing It or Not? Interpretable Vision-aware Latent Steering to Mitigate Object Hallucinations Detecting and preventing hallucinations in large vision language models

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:45:36.895414Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-07T14:45:29.202872Z digest=sha256:c8126564d119bd5f3eec7609ac166f4f1ed4bb7b4470d2f3b7bc668a9175711a

Observation 125f43ae-65de-4b25-be32-532fb592a6ec · outbound

This paper cites Aligning Large Multimodal Models with Factually Augmented RLHF.

Seeing It or Not? Interpretable Vision-aware Latent Steering to Mitigate Object Hallucinations Aligning Large Multimodal Models with Factually Augmented RLHF

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-07T14:45:29.287602Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:45:29.287602Z digest=sha256:97bc1e3c517c59b77ad7da51aaa4dff822491894ec6010f6441060de12fcda6a

Observation 85fa9ad1-866f-4ebb-b530-46f4fd57c7f6 · outbound

This paper cites Dress: Instructing large vision-language models to align and interact with humans via natural language feedback.

Seeing It or Not? Interpretable Vision-aware Latent Steering to Mitigate Object Hallucinations Dress: Instructing large vision-language models to align and interact with humans via natural language feedback

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:45:36.712636Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-07T14:45:29.363088Z digest=sha256:a8ae10cbee79d0eeb7f95aa5a9f12af1e4dbe0ee05838f92f016007921d67bb6

Observation 7f9b0304-50f3-4f2e-b0ec-d8cabe019404 · outbound

This paper cites Woodpecker: Hallucination Correction for Multimodal Large Language Models.

Seeing It or Not? Interpretable Vision-aware Latent Steering to Mitigate Object Hallucinations Woodpecker: Hallucination Correction for Multimodal Large Language Models

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-07T14:45:29.439784Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:45:29.439784Z digest=sha256:b5125245dbffc9ec41a4c6ee4f4e6a60232db8e0ee0cf0f0873487c28a61a515

Observation a0f958ef-718f-482d-b0f9-39b5dd6e23b3 · outbound

This paper cites Mitigating object hallucination in large vision-language models via image-grounded guidance.

Seeing It or Not? Interpretable Vision-aware Latent Steering to Mitigate Object Hallucinations Mitigating object hallucination in large vision-language models via image-grounded guidance

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:45:36.500425Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-07T14:45:29.524121Z digest=sha256:1aec28a6ea87396882ad71401f9fc69081203179086376b759d72a3e0c615d40

Observation ff3e8d54-95f9-4d80-b404-ff620b82690d · outbound

This paper cites Paying more attention to image: A training-free method for alleviating hallucination in lvlms.

Seeing It or Not? Interpretable Vision-aware Latent Steering to Mitigate Object Hallucinations Paying more attention to image: A training-free method for alleviating hallucination in lvlms

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:45:36.337308Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-07T14:45:29.593024Z digest=sha256:86e65dbad232be81e95533022a3a9ee1e6746af6931d24c2bdda90bc87cf03c0

Observation b6698032-96ff-40d5-a3aa-b68b7600a43a · outbound

This paper cites IBD: Alleviating Hallucinations in Large Vision-Language Models via Image-Biased Decoding.

Seeing It or Not? Interpretable Vision-aware Latent Steering to Mitigate Object Hallucinations IBD: Alleviating Hallucinations in Large Vision-Language Models via Image-Biased Decoding

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-07T14:45:29.672408Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:45:29.672408Z digest=sha256:6f4297f91a2267c1595dc51e3ebaa2c4f53afba5afd935d6ada6204a356afd95

Observation a24c3d4f-df58-4cf0-bcd3-cb50f4e2da2d · outbound

This paper cites Opera: Alleviating hallucination in multi-modal large language models via over-trust penalty and retrospection-allocation.

Seeing It or Not? Interpretable Vision-aware Latent Steering to Mitigate Object Hallucinations Opera: Alleviating hallucination in multi-modal large language models via over-trust penalty and retrospection-allocation

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:45:36.102530Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-07T14:45:29.758823Z digest=sha256:aba58059577d4cbb2afb283ae8c82a712ca88ee17c0507bdc6ae7a3f5e6b024e

Observation 77a176db-9161-4150-9360-879d127009a8 · outbound

This paper cites Multi-modal hallucination control by visual information grounding.

Seeing It or Not? Interpretable Vision-aware Latent Steering to Mitigate Object Hallucinations Multi-modal hallucination control by visual information grounding

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:45:35.874270Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-07T14:45:29.857326Z digest=sha256:567aa99f83e333491ae1e266c193251af46176bf5a03c467e461e69aa63713c9

Observation a06fcb51-2132-4696-a792-026aa2d63c38 · outbound

This paper cites AlphaEdit: Null-Space Constrained Knowledge Editing for Language Models.

Seeing It or Not? Interpretable Vision-aware Latent Steering to Mitigate Object Hallucinations AlphaEdit: Null-Space Constrained Knowledge Editing for Language Models

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-07T14:45:29.943737Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:45:29.943737Z digest=sha256:59eb94deccf31ce0ed2271da097a35f0404b99cf481492633124d27f0d6d7280

Observation b2eb7736-2804-407e-9555-9576ea099f9f · outbound

This paper cites Mitigating Object Hallucinations in Large Vision-Language Models with Assembly of Global and Local Attention.

Seeing It or Not? Interpretable Vision-aware Latent Steering to Mitigate Object Hallucinations Mitigating Object Hallucinations in Large Vision-Language Models with Assembly of Global and Local Attention

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-07T14:45:30.042461Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:45:30.042461Z digest=sha256:5b7182aa5d86e779151f06d1402d7a0c5b976fecc7c653a28a379ad1f9aa9551

Observation b4fefc28-d1a6-4b60-ab67-38a9e25413b6 · outbound

This paper cites Grad-cam: Visual explanations from deep networks via gradient-based localization.

Seeing It or Not? Interpretable Vision-aware Latent Steering to Mitigate Object Hallucinations Grad-cam: Visual explanations from deep networks via gradient-based localization

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:45:35.675866Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-07T14:45:30.110923Z digest=sha256:d27255007cc3b6acfdea4af2de370e6b3ee17a0eca487600eaa80fb5fae7e144

Observation 847e83b9-e414-4ccd-9632-e2aecb5eb7ab · outbound

This paper cites Grad-cam++: Generalized gradient-based visual explanations for deep convolutional networks.

Seeing It or Not? Interpretable Vision-aware Latent Steering to Mitigate Object Hallucinations Grad-cam++: Generalized gradient-based visual explanations for deep convolutional networks

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:45:35.484548Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-07T14:45:30.219446Z digest=sha256:09306e0b2102c28d99bae70ab12f9910d35a70833dda703a0c62be2852d3ebca

Observation 3284945c-49ca-4b6d-9391-ef63178ae38a · outbound

This paper cites Generic attention-model explainability for interpreting bi-modal and encoder-decoder transformers.

Seeing It or Not? Interpretable Vision-aware Latent Steering to Mitigate Object Hallucinations Generic attention-model explainability for interpreting bi-modal and encoder-decoder transformers

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:45:35.230180Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-07T14:45:30.320579Z digest=sha256:8a6e2388a679c702f8bbdb89e07b8d2bf273d429a63f376b4ede3603f9e8f74a

Observation 8272d023-731e-4d73-b974-2645efdb56d7 · outbound

This paper cites Transformer interpretability beyond attention visualization.

Seeing It or Not? Interpretable Vision-aware Latent Steering to Mitigate Object Hallucinations Transformer interpretability beyond attention visualization

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:45:35.077456Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-07T14:45:30.399480Z digest=sha256:23ddafc8a60db092174c21104eedfbe23ce6a8673dfecfae239314ad784a1dc3

Observation 82447b30-d45b-4a50-a20a-83d9ff7a2806 · outbound

This paper cites Vl- interpret: An interactive visualization tool for interpreting vision-language transformers.

Seeing It or Not? Interpretable Vision-aware Latent Steering to Mitigate Object Hallucinations Vl- interpret: An interactive visualization tool for interpreting vision-language transformers

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:45:34.838376Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-07T14:45:30.463305Z digest=sha256:efc333e420dcea2bf8e25fc7c2ac2d1e119e4e4bffd5ec952a5f834c64ba6bc1

Observation 2c82cd78-bd42-4df1-8c17-3405ca1d6741 · outbound

This paper cites FastRM: An efficient and automatic explainability framework for multimodal generative models.

Seeing It or Not? Interpretable Vision-aware Latent Steering to Mitigate Object Hallucinations FastRM: An efficient and automatic explainability framework for multimodal generative models

Reference 48

Resolution
verified exact
local_arxiv, observed 2026-08-07T14:45:32.849467Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-07T14:45:30.527268Z digest=sha256:aa4865d3505d196e4dfb45251156b409a13d16c4845168ec04140b8fd78aa4e1

Observation 3519c456-6106-4f20-b0df-74c2d34808f7 · outbound

This paper cites Explaining Multi-modal Large Language Models by Analyzing their Vision Perception.

Seeing It or Not? Interpretable Vision-aware Latent Steering to Mitigate Object Hallucinations Explaining Multi-modal Large Language Models by Analyzing their Vision Perception

Reference 49

Resolution
verified exact
local_arxiv, observed 2026-08-07T14:45:32.614357Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-07T14:45:30.624312Z digest=sha256:439b341da398b4941fa0995325b560227f9bef86d5d2da1e0bfe58adc396762d

Observation 7089ab6c-d0a4-4247-8caf-b136bf755174 · outbound

This paper cites From Redundancy to Relevance: Information Flow in LVLMs Across Reasoning Tasks.

Seeing It or Not? Interpretable Vision-aware Latent Steering to Mitigate Object Hallucinations From Redundancy to Relevance: Information Flow in LVLMs Across Reasoning Tasks

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-07T14:45:30.728861Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:45:30.728861Z digest=sha256:1f493fba657f6579806fb8e87b21003716e0a4dcf5bfdf06fe689e7f39ace9e2

Observation bc37ab4e-748e-45aa-aa7f-6b302367a133 · outbound

This paper cites Finding and Editing Multi-Modal Neurons in Pre-Trained Transformers.

Seeing It or Not? Interpretable Vision-aware Latent Steering to Mitigate Object Hallucinations Finding and Editing Multi-Modal Neurons in Pre-Trained Transformers

Reference 51

Resolution
verified exact
local_arxiv, observed 2026-08-07T14:45:32.303375Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-07T14:45:30.773041Z digest=sha256:9d5d2590af9e727a7e3ddbe958ef48cc2d22e870df9bb9ab4cd656a995d44380

Observation 6ccb7e24-eaa5-4926-accd-cfd30a528366 · outbound

This paper cites AMBER: An LLM-free Multi-dimensional Benchmark for MLLMs Hallucination Evaluation.

Seeing It or Not? Interpretable Vision-aware Latent Steering to Mitigate Object Hallucinations AMBER: An LLM-free Multi-dimensional Benchmark for MLLMs Hallucination Evaluation

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-07T14:45:30.862604Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:45:30.862604Z digest=sha256:24e0f8300edea4a893b37f19c04a4039ebd0838d42cee59004f5d365736427f7

Observation 1889192c-40e8-4d30-96c2-56425dc377f8 · outbound

This paper cites Evaluating object hallucination in large vision-language models.

Seeing It or Not? Interpretable Vision-aware Latent Steering to Mitigate Object Hallucinations Evaluating object hallucination in large vision-language models

Reference 53

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:45:34.654011Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-07T14:45:30.949942Z digest=sha256:5eb3e0e11507f649d914888b47328ab095dbebd4e554235f75ecc8ef4037ed23

Observation 3e5446db-90a3-421c-a606-095f94019fde · outbound

This paper cites Aligning large multimodal models with factually augmented rlhf.

Seeing It or Not? Interpretable Vision-aware Latent Steering to Mitigate Object Hallucinations Aligning large multimodal models with factually augmented rlhf

Reference 54

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:45:34.471205Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-07T14:45:31.024274Z digest=sha256:36c62bc4c425ed3523d7595ef0e15760c277e332a1807471cf8e26f12acd43ee

Observation 707ed996-fdbb-4892-94d2-a0dcd63b2e9c · outbound

This paper cites Videomae: Masked autoencoders are data-efficient learners for self-supervised video pre-training.Adv.

Seeing It or Not? Interpretable Vision-aware Latent Steering to Mitigate Object Hallucinations Videomae: Masked autoencoders are data-efficient learners for self-supervised video pre-training.Adv

Reference 55

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:45:34.299454Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-07T14:45:31.070353Z digest=sha256:efbe8607df203942f59a0809b52773e218e18b4e338048b77d83bc196961cf5f

Observation 6f82f695-ebfd-409f-8796-bc9795f604ae · outbound

This paper cites MME: A Comprehensive Evaluation Benchmark for Multimodal Large Language Models.

Seeing It or Not? Interpretable Vision-aware Latent Steering to Mitigate Object Hallucinations MME: A Comprehensive Evaluation Benchmark for Multimodal Large Language Models

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-07T14:45:31.150819Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:45:31.150819Z digest=sha256:1e999ef1aad12e9beff1a60f2764ea6a869b685c0276dd28abf00e66ca7319f9

Observation 4fe76543-d92b-4d9d-af14-a3bec28a87d2 · outbound

This paper cites Gqa: A new dataset for real-world visual reasoning and compositional question answering.

Seeing It or Not? Interpretable Vision-aware Latent Steering to Mitigate Object Hallucinations Gqa: A new dataset for real-world visual reasoning and compositional question answering

Reference 57

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:45:34.114986Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-07T14:45:31.203608Z digest=sha256:29cc5d0e13cf4b796cf05e03ee285b9b3f9c742b006cdf85ff11a379f99750dd

Observation ebdf2516-b466-4eaa-ace5-f91f494677cb · outbound

This paper cites Beam Search Strategies for Neural Machine Translation.

Seeing It or Not? Interpretable Vision-aware Latent Steering to Mitigate Object Hallucinations Beam Search Strategies for Neural Machine Translation

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-07T14:45:31.289807Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:45:31.289807Z digest=sha256:be8d11863135eab126960d94e127c4fd337a5a150a6bba388fbe6db92075312c

Observation 4de1a212-ef38-42f1-8a30-cf3c92771f3c · outbound

This paper cites DoLa: Decoding by Contrasting Layers Improves Factuality in Large Language Models.

Seeing It or Not? Interpretable Vision-aware Latent Steering to Mitigate Object Hallucinations DoLa: Decoding by Contrasting Layers Improves Factuality in Large Language Models

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-07T14:45:31.400584Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:45:31.400584Z digest=sha256:0e49c44b54689dbd68f7b81b30017a1c25ee7f83146bd440d73cceb342a7c894

Observation f8bab591-d094-47bc-931f-2d936769cd25 · outbound

This paper cites Blended diffusion for text-driven editing of natural images.

Seeing It or Not? Interpretable Vision-aware Latent Steering to Mitigate Object Hallucinations Blended diffusion for text-driven editing of natural images

Reference 60

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:45:33.960642Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-07T14:45:31.506177Z digest=sha256:91f33bf7da7762878665f2416fdaee8e8c2474508198835a9d3e598f1ea8109a

Observation b87fca78-67d1-42ec-9d67-d2cc6903bcab · outbound

This paper cites Unveiling typographic deceptions: Insights of the typographic vulnerability in large vision-language models.

Seeing It or Not? Interpretable Vision-aware Latent Steering to Mitigate Object Hallucinations Unveiling typographic deceptions: Insights of the typographic vulnerability in large vision-language models

Reference 61

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:45:33.708879Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-07T14:45:31.619307Z digest=sha256:400326d9457408acb4f78d60ba90231cb94279ea3f0e7fdfd6854d971a72fec1

Observation 2e495d30-7f69-4a63-92fe-e137a6360e93 · outbound

This paper cites Real-time single image and video super-resolution using an efficient sub-pixel convolutional neural network.

Seeing It or Not? Interpretable Vision-aware Latent Steering to Mitigate Object Hallucinations Real-time single image and video super-resolution using an efficient sub-pixel convolutional neural network

Reference 62

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:45:33.567854Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-07T14:45:31.774717Z digest=sha256:1d7dd22d550d9a35f761f7c6845517e8a846e2a70037d31d2e615815288a88d0

Observation f7264abe-3342-48d1-8953-f240882fb7ca · outbound

This paper cites Towards interpreting visual information processing in vision-language models.

Seeing It or Not? Interpretable Vision-aware Latent Steering to Mitigate Object Hallucinations Towards interpreting visual information processing in vision-language models

Reference 63

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:45:33.440316Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-07T14:45:31.888895Z digest=sha256:691dd75e3ac02df88c19438f61c84cf244937250d4ae3c93571dac6f80a7c4ba

Observation bf173a22-c628-4ba6-8068-16487a728e86 · outbound

This paper cites GPT-4 Technical Report.

Seeing It or Not? Interpretable Vision-aware Latent Steering to Mitigate Object Hallucinations GPT-4 Technical Report

Reference 64

Resolution
unresolved
no resolver link, observed 2026-08-07T14:45:31.991900Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:45:31.991900Z digest=sha256:0d2a27a871187d8c54ed8fcb7f34a69b54ea59fbd30ce0cb74b9c851612b98f4

Observation e4e60157-5b8f-4138-bf26-7e5b1f2d07bd · outbound

This paper cites Eyes Wide Shut? Exploring the Visual Shortcomings of Multimodal LLMs.

Seeing It or Not? Interpretable Vision-aware Latent Steering to Mitigate Object Hallucinations Eyes Wide Shut? Exploring the Visual Shortcomings of Multimodal LLMs

Reference 65

Resolution
malformed identifier
no resolver link, observed 2026-08-07T14:45:32.061891Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:45:32.061891Z digest=sha256:bc672e54a0e3b7cc46dd26060cc0f1b94f03d8f06650fe0cc81989a742aa058a

Pith citing papers

No inbound Pith citation observations are available.