Pith. sign in

Paper Citation Record · LEDGER

Object Hallucination in Image Captioning

As of 5 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 46 inbound Pith citation observations for arXiv:1809.02156.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
1809.02156 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 46 of 46 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-05T06:32:48.257954+00:00

measured 46 of 46 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-04T08:14:36.635360Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-07-03T19:28:52.843969Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 697296b5-2ba4-49ac-ac55-941eb97f7095 · inbound

MiniGPT-4: Enhancing Vision-Language Understanding with Advanced Large Language Models cites this paper.

MiniGPT-4: Enhancing Vision-Language Understanding with Advanced Large Language Models Object Hallucination in Image Captioning

Reference 14

Resolution
verified exact
arxiv_id, observed 2026-05-10T20:37:01.921715Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T20:37:01.617345Z digest=sha256:55ce2da83f98c09a8d1925b11d24b2756e7d88b770b0d4d002cc4282922782bc

Observation 1a3b89bf-c79a-4ed3-a493-bde664c9b593 · inbound

Mitigating Hallucination in Large Multi-Modal Models via Robust Instruction Tuning cites this paper.

Mitigating Hallucination in Large Multi-Modal Models via Robust Instruction Tuning Object Hallucination in Image Captioning

Reference 17

Resolution
verified exact
local_arxiv, observed 2026-05-14T17:34:56.969543Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-14T17:34:56.836034Z digest=sha256:434fbc1f1945c751fcb811dcbf54072e4274647bc80b4cd0502e892be7b699d4

Observation afd2db2e-72be-48d2-9a91-4e7baa3c2bb6 · inbound

A Survey of Hallucination in Large Foundation Models cites this paper.

A Survey of Hallucination in Large Foundation Models Object Hallucination in Image Captioning

Reference 141

Resolution
verified exact
local_arxiv, observed 2026-05-16T15:21:00.977622Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-16T15:21:00.778049Z digest=sha256:db318d8cee0636387945bc27f7069da3cd7c257e3f8b4b1d83f37d82a634415b

Observation 6ae1d947-e68b-4808-bfcf-492bbec5c350 · inbound

Aligning Large Multimodal Models with Factually Augmented RLHF cites this paper.

Aligning Large Multimodal Models with Factually Augmented RLHF Object Hallucination in Image Captioning

Reference 25

Resolution
verified exact
local_arxiv, observed 2026-05-15T17:58:17.901268Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-15T17:58:17.699042Z digest=sha256:e3d803db56562fe9921c8d5f55f34c3857fcf760d5c19ff770f3dd675bc9c4de

Observation 37e9a718-5127-4d20-a870-4188b80be376 · inbound

MiniGPT-v2: large language model as a unified interface for vision-language multi-task learning cites this paper.

MiniGPT-v2: large language model as a unified interface for vision-language multi-task learning Object Hallucination in Image Captioning

Reference 40

Resolution
verified exact
local_arxiv, observed 2026-05-16T07:13:08.939153Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-16T07:13:08.867745Z digest=sha256:83d4c917be1fad61e87bf8a511d82a6ce5dc3f5e635dacb910d8da43936f5145

Observation b2ecb0dd-577a-4dee-a739-ebb90b134473 · inbound

AMBER: An LLM-free Multi-dimensional Benchmark for MLLMs Hallucination Evaluation cites this paper.

AMBER: An LLM-free Multi-dimensional Benchmark for MLLMs Hallucination Evaluation Object Hallucination in Image Captioning

Reference 9

Resolution
metadata mismatch
local_arxiv, observed 2026-05-16T06:44:30.280751Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-16T06:44:30.244832Z digest=sha256:89683d5483d3c018f5fad6941a8d74c4a17182ec02ed7e6d766f07a8b2e1bed6

Observation 23a75bd6-065b-46df-bfdc-6a01dc2345b5 · inbound

Agent AI: Surveying the Horizons of Multimodal Interaction cites this paper.

Agent AI: Surveying the Horizons of Multimodal Interaction Object Hallucination in Image Captioning

Reference 190

Resolution
metadata mismatch
local_arxiv, observed 2026-05-18T14:25:59.598259Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-18T14:25:58.876978Z digest=sha256:ec94c574fde5fee51539c8c811259f33f504a23718687123bb83a5dcd2f7488b

Observation 912d49ae-3342-4872-9dd3-1ea18b68a802 · inbound

Aligning Modalities in Vision Large Language Models via Preference Fine-tuning cites this paper.

Aligning Modalities in Vision Large Language Models via Preference Fine-tuning Object Hallucination in Image Captioning

Reference 170

Resolution
metadata mismatch
local_arxiv, observed 2026-05-17T10:58:53.453465Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-17T10:58:53.215887Z digest=sha256:1b760d96fee53434c7ba46c9b6962adb28290a6933875e4c39d93d83e9056cd4

Observation 464124d0-21a7-4035-89c8-27f12517beba · inbound

Hallucination of Multimodal Large Language Models: A Survey cites this paper.

Hallucination of Multimodal Large Language Models: A Survey Object Hallucination in Image Captioning

Reference 141

Resolution
verified exact
arxiv_id, observed 2026-05-11T12:33:32.953529Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-11T12:33:32.631346Z digest=sha256:768c97392cf1c53f6621ba50bd404b9778f43268176029b75fa1615d0d3b3b01

Observation 07f0b878-a985-4302-aeb6-699d9f49acfb · inbound

Detecting and Evaluating Medical Hallucinations in Large Vision Language Models cites this paper.

Detecting and Evaluating Medical Hallucinations in Large Vision Language Models Object Hallucination in Image Captioning

Reference 19

Resolution
verified exact
local_arxiv, observed 2026-05-23T23:58:39.544638Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-23T23:55:57.103971Z digest=sha256:54b4a0d996eaea5a6ae32fc48c4dcbc0f51d810b02697117d75d8442e62cf9f9

Observation 85efcd41-1440-4f05-bd22-275b9eada7f5 · inbound

MiniCPM-V: A GPT-4V Level MLLM on Your Phone cites this paper.

MiniCPM-V: A GPT-4V Level MLLM on Your Phone Object Hallucination in Image Captioning

Reference 85

Resolution
verified exact
arxiv_id, observed 2026-05-10T21:07:31.834351Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T21:07:31.387726Z digest=sha256:afa32865d35e410cf3bc39586485460fb178a45a24f6102a1c8ee23604f9c2ca

Observation 5e4db803-7756-4878-9e48-9c4aebfbdc3d · inbound

Enhancing the Reasoning Ability of Multimodal Large Language Models via Mixed Preference Optimization cites this paper.

Enhancing the Reasoning Ability of Multimodal Large Language Models via Mixed Preference Optimization Object Hallucination in Image Captioning

Reference 82

Resolution
verified exact
local_arxiv, observed 2026-05-16T09:16:17.417120Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-16T09:16:17.150383Z digest=sha256:498ab232dc5ec63364a72df85cd5a8d82851afa4f5f9c883628985c265159e25

Observation 156fd32f-0218-4bac-bc0f-a2a7edf7aef3 · inbound

Mitigating Hallucination in Large Vision-Language Models via Adaptive Attention Calibration cites this paper.

Mitigating Hallucination in Large Vision-Language Models via Adaptive Attention Calibration Object Hallucination in Image Captioning

Reference 15

Resolution
metadata mismatch
local_arxiv, observed 2026-05-19T12:52:18.055019Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-19T12:48:44.324236Z digest=sha256:b9aeaab7a37379fdac08689a8eab1c59f33b2408ca101c78430ea5b92e160c18

Observation 363de7a5-aaa4-47de-a927-971a141fafda · inbound

ZINA: Multimodal Fine-grained Hallucination Detection and Editing cites this paper.

ZINA: Multimodal Fine-grained Hallucination Detection and Editing Object Hallucination in Image Captioning

Reference 3

Resolution
metadata mismatch
local_arxiv, observed 2026-05-19T10:12:14.611481Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-19T10:07:18.162776Z digest=sha256:563cf015e2211b8565a122c3dd5f9e36800db49e3a045a6a6b1565cf06cfd07b

Observation 90974ff9-9fd6-4645-9c8f-5fa7cc58ff8c · inbound

Mitigating Object Hallucinations via Sentence-Level Early Intervention cites this paper.

Mitigating Object Hallucinations via Sentence-Level Early Intervention Object Hallucination in Image Captioning

Reference 55

Resolution
verified exact
local_arxiv, observed 2026-05-25T08:35:32.589805Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-25T08:31:24.173135Z digest=sha256:a972b25d054db1e1dde4a1c17062071a880785fbe8bf519d68d53147e8a64a80

Observation 5ef232f3-ed38-462a-8175-99bae75ff279 · inbound

Capturing Gaze Shifts for Guidance: Cross-Modal Fusion Enhancement for VLM Hallucination Mitigation cites this paper.

Capturing Gaze Shifts for Guidance: Cross-Modal Fusion Enhancement for VLM Hallucination Mitigation Object Hallucination in Image Captioning

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-04T08:14:36.635360Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T08:14:36.635360Z digest=sha256:ff65c1546fe91bfc1f30aeb945b78f58099b62ec2804e2997662050afc5b34e1

Observation 616a1539-8c60-4420-9c04-635cc96b5333 · inbound

Adaptive Residual-Update Steering for Low-Overhead Hallucination Mitigation in Large Vision Language Models cites this paper.

Adaptive Residual-Update Steering for Low-Overhead Hallucination Mitigation in Large Vision Language Models Object Hallucination in Image Captioning

Reference 17

Resolution
verified exact
local_arxiv, observed 2026-05-21T18:50:30.057037Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-21T18:47:28.768554Z digest=sha256:4398d4c8b2d1db5afb55e20a8ba91153a6c8eaa4ec5e59a1bb489085450f1f5b

Observation 8a5bc30f-73c9-4660-aec1-ded9fe32b0c4 · inbound

A Survey on Evaluating Quality and Trustworthiness in LLM-Generated Data cites this paper.

A Survey on Evaluating Quality and Trustworthiness in LLM-Generated Data Object Hallucination in Image Captioning

Reference 182

Resolution
unresolved
no resolver link, observed 2026-08-03T08:15:28.752689Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-03T08:15:28.752689Z digest=sha256:6b3820b95a0a415b638237db445c9198bd4759bf9a1baba1528ea1b465e494b4

Observation 3aa2dc33-9505-4445-8ea2-77467f9fc032 · inbound

MACD: Model-Aware Contrastive Decoding via Counterfactual Data cites this paper.

MACD: Model-Aware Contrastive Decoding via Counterfactual Data Object Hallucination in Image Captioning

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-03T05:37:28.544825Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T05:37:28.544825Z digest=sha256:3aab058a04a0933e4b7ddc3d7876678cea02d46a99a03193d08a965183a4af0e

Observation c6284b07-ff09-4ad0-8f25-c98f71f3226d · inbound

Contextualized Visual Personalization in Vision-Language Models cites this paper.

Contextualized Visual Personalization in Vision-Language Models Object Hallucination in Image Captioning

Reference 1

Resolution
metadata mismatch
local_arxiv, observed 2026-05-16T08:20:45.961908Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-16T08:20:22.025379Z digest=sha256:010a0090eac23a2757e9f3c9e1ca0ccb05c049736979bfbdcdba1285d859c608

Observation 9b9799f1-6fae-4654-98a6-46fb04276da0 · inbound

Contextualized Visual Personalization in Vision-Language Models cites this paper.

Contextualized Visual Personalization in Vision-Language Models Object Hallucination in Image Captioning

Reference 1

Resolution
metadata mismatch
local_arxiv, observed 2026-05-21T14:20:13.765484Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-21T14:17:12.540477Z digest=sha256:01fd3352674c0353d6f76008888ea3b8fd8baa7d2f0b07e0a9916bb766c98469

Observation 419e7015-7d4e-42af-b322-6c472d40db17 · inbound

Revis: Sparse Latent Steering to Mitigate Object Hallucination in Large Vision-Language Models cites this paper.

Revis: Sparse Latent Steering to Mitigate Object Hallucination in Large Vision-Language Models Object Hallucination in Image Captioning

Reference 6

Resolution
metadata mismatch
local_arxiv, observed 2026-05-16T05:07:20.503881Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-16T05:06:24.973439Z digest=sha256:ca52c25ed6a304db2bca01b9ec9c83d781d96cb466363f6a90e8089e457ffc5f

Observation 96674d73-8402-403a-b229-bcca37873af5 · inbound

Spotlight and Shadow: Attention-Guided Dual-Anchor Introspective Decoding for MLLM Hallucination Mitigation cites this paper.

Spotlight and Shadow: Attention-Guided Dual-Anchor Introspective Decoding for MLLM Hallucination Mitigation Object Hallucination in Image Captioning

Reference 26

Resolution
verified exact
arxiv_id, observed 2026-05-11T07:35:57.420240Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-10T17:07:09.330134Z digest=sha256:b0601d946f5839b6587a3b00cec88d752cca5c7a9fcb559e08d35589ee49fc06

Observation 059541a9-1080-4e56-82a0-9e0d99d23708 · inbound

Mitigating Multimodal Hallucination via Phase-wise Self-reward cites this paper.

Mitigating Multimodal Hallucination via Phase-wise Self-reward Object Hallucination in Image Captioning

Reference 39

Resolution
verified exact
arxiv_id, observed 2026-05-10T09:43:49.330311Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T05:10:45.144421Z digest=sha256:4cf48c5713231c2a959210ed2b9c42a991bd1d9646143611335f9c26b8d43972

Observation e2d14852-df1f-4934-9cc3-b0696f1d6f4a · inbound

Mitigating Hallucinations in Large Vision-Language Models without Performance Degradation cites this paper.

Mitigating Hallucinations in Large Vision-Language Models without Performance Degradation Object Hallucination in Image Captioning

Reference 137

Resolution
metadata mismatch
arxiv_id, observed 2026-05-10T00:34:47.764596Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-10T00:33:39.960170Z digest=sha256:b6087217cb39925c849175128882684a7438e6465c9e2f6ec3880d5a6d27ce49

Observation 5b057bdf-cdab-4fb8-8620-8d6c562a406e · inbound

Prefill-Time Intervention for Mitigating Hallucination in Large Vision-Language Models cites this paper.

Prefill-Time Intervention for Mitigating Hallucination in Large Vision-Language Models Object Hallucination in Image Captioning

Reference 33

Resolution
verified exact
arxiv_id, observed 2026-05-11T23:31:16.328906Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-07T16:50:26.044591Z digest=sha256:45306ae913dccf7b2f3ec523628c0df536862d85df1d8a3a50756f09a972a6ea

Observation 612ce53a-bc2f-4fd5-880f-5cb64366a721 · inbound

Online Self-Calibration Against Hallucination in Vision-Language Models cites this paper.

Online Self-Calibration Against Hallucination in Vision-Language Models Object Hallucination in Image Captioning

Reference 23

Resolution
verified exact
arxiv_id, observed 2026-05-11T15:16:10.535343Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-09T20:20:27.931679Z digest=sha256:daeeaf44cf9aab1ae494bdf9890ff8a063a55c3a9240999d71fe51118bf36f3e

Observation 66039d67-07d8-4d17-9015-699d0f2547ff · inbound

CAST: Mitigating Object Hallucination in Large Vision-Language Models via Caption-Guided Visual Attention Steering cites this paper.

CAST: Mitigating Object Hallucination in Large Vision-Language Models via Caption-Guided Visual Attention Steering Object Hallucination in Image Captioning

Reference 40

Resolution
metadata mismatch
arxiv_id, observed 2026-05-09T06:50:40.149814Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-08T18:05:38.705480Z digest=sha256:c56c4f7d88778cabd9dd6bfb973b654dc63c44b56ec14065fc1a7c3e1949c87c

Observation e18c536d-7d25-4f01-8667-7633e712a847 · inbound

Mitigating Action-Relation Hallucinations in LVLMs via Relation-aware Visual Enhancement cites this paper.

Mitigating Action-Relation Hallucinations in LVLMs via Relation-aware Visual Enhancement Object Hallucination in Image Captioning

Reference 42

Resolution
metadata mismatch
arxiv_id, observed 2026-05-13T05:47:21.426404Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-13T05:44:05.093926Z digest=sha256:8522fad51e1e5b1e141e5618254bb1159489a8e253a75b17aa136d802367136b

Observation 5100c479-7e96-4acd-8b2f-2ff752075241 · inbound

From Failure to Feedback: Group Revision Unlocks Hard Cases in Object-Level Grounding cites this paper.

From Failure to Feedback: Group Revision Unlocks Hard Cases in Object-Level Grounding Object Hallucination in Image Captioning

Reference 67

Resolution
verified exact
local_arxiv, observed 2026-05-20T18:43:38.836035Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-20T18:39:11.904941Z digest=sha256:341c8a73d9dd3b56a4fc2cfda571b7df80e3e03cc91ac24023e04ae75b55b6ca

Observation a2bb9b18-8b74-41c2-b1ca-37b5334da984 · inbound

How do Humans Process AI-generated Hallucination Contents: a Neuroimaging Study cites this paper.

How do Humans Process AI-generated Hallucination Contents: a Neuroimaging Study Object Hallucination in Image Captioning

Reference 8

Resolution
metadata mismatch
local_arxiv, observed 2026-05-19T20:37:45.365630Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-19T20:33:17.097346Z digest=sha256:7bf70d94facf070a25386b0530bf17b407de851a0e4b6d365bb08bbee6c08ad5

Observation 0da25dd2-1054-4d29-8227-e8de23cd9ef1 · inbound

How do Humans Process AI-generated Hallucination Contents: a Neuroimaging Study cites this paper.

How do Humans Process AI-generated Hallucination Contents: a Neuroimaging Study Object Hallucination in Image Captioning

Reference 10

Resolution
metadata mismatch
local_arxiv, observed 2026-06-30T19:15:00.645349Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-06-30T19:10:49.286658Z digest=sha256:ef19a69cd4e1cd34beee9bac374d57dde10954f06ede94343215ca1b56bc9c84

Observation 82982a83-c4f9-44b2-9b15-8c48268b3765 · inbound

Focus-then-Context: Subject-Centric Progressive Visual Token Reduction for Vision-Language Models cites this paper.

Focus-then-Context: Subject-Centric Progressive Visual Token Reduction for Vision-Language Models Object Hallucination in Image Captioning

Reference 42

Resolution
metadata mismatch
local_arxiv, observed 2026-05-21T05:23:58.590578Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-21T05:20:55.448430Z digest=sha256:443498cf6c1f32ed5922cf937a04754ac960bf63ca88e61902c0852f2101d54c

Observation f2888206-de41-456a-a560-92d399075141 · inbound

Rethinking Visual Neglect: Steering via Context-Preference for MLLM Hallucination Mitigation cites this paper.

Rethinking Visual Neglect: Steering via Context-Preference for MLLM Hallucination Mitigation Object Hallucination in Image Captioning

Reference 21

Resolution
verified exact
local_arxiv, observed 2026-06-29T13:33:28.033520Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-06-29T13:29:33.614965Z digest=sha256:c8193c6f97165fa2b0b60e55debf6f222241901bcf374d9456ab7ef1132d5b3a

Observation cd130a9f-bfc6-48b2-94b0-9e8c3f14eb6d · inbound

What Makes LVLMs Hallucinate Less? Unveiling the Architectural Factors Behind Hallucination Robustness cites this paper.

What Makes LVLMs Hallucinate Less? Unveiling the Architectural Factors Behind Hallucination Robustness Object Hallucination in Image Captioning

Reference 39

Resolution
verified exact
local_arxiv, observed 2026-06-28T23:22:46.588511Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-06-28T23:21:50.254449Z digest=sha256:2ae0d1de3d1078b717a7dd644a542445d212646de911ba323d949f9d6ff2e0a4

Observation defdc462-4b8c-4e0b-89aa-4fa6c166045d · inbound

Hyper-ICL: Attention Calibration with Hyperbolic Anchor Distillation for Multimodal In-Context Learning cites this paper.

Hyper-ICL: Attention Calibration with Hyperbolic Anchor Distillation for Multimodal In-Context Learning Object Hallucination in Image Captioning

Reference 7

Resolution
metadata mismatch
local_arxiv, observed 2026-07-02T06:36:43.897031Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-06-28T07:23:19.428348Z digest=sha256:dbd6c8c6b9d260b0c99d97713e71d6a565301285cc3f33207007965ff2a1981a

Observation 0ebce830-91cf-4b04-89b4-9986d799d095 · inbound

Steer Where It Matters: Token-Level Visual-Sensitivity Steering for LVLMs Hallucination Mitigation cites this paper.

Steer Where It Matters: Token-Level Visual-Sensitivity Steering for LVLMs Hallucination Mitigation Object Hallucination in Image Captioning

Reference 13

Resolution
metadata mismatch
local_arxiv, observed 2026-07-02T03:06:29.659703Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-06-28T10:22:40.055153Z digest=sha256:790573db4ec47973213fc16fbd2da1f8c2449684fac75ba390e010aa3f9092be

Observation c2a7b65c-f690-4133-ac7c-94bd81e40a74 · inbound

Task-Aware Structured Memory for Dynamic Multi-modal In-Context Learning cites this paper.

Task-Aware Structured Memory for Dynamic Multi-modal In-Context Learning Object Hallucination in Image Captioning

Reference 56

Resolution
metadata mismatch
local_arxiv, observed 2026-07-03T09:07:48.423984Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-06-27T10:28:11.440915Z digest=sha256:56b010942dc1ac0b226891eb0064b3fc41c5f3fa86a3ac7131d20014f0a08423

Observation ef07b4d4-3c01-40db-b381-f8e96953689e · inbound

See First, Answer Later: Visual Evidence Pre-Alignment via Sufficiency-Driven RL cites this paper.

See First, Answer Later: Visual Evidence Pre-Alignment via Sufficiency-Driven RL Object Hallucination in Image Captioning

Reference 32

Resolution
metadata mismatch
local_arxiv, observed 2026-07-03T19:28:52.845331Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-06-27T01:44:24.586195Z digest=sha256:e26d2eaa875049316578d41ccd923eaa236c9275f04812509186f6c5991d47cd

Observation 7ad6c26b-1e76-4801-ac32-e683167afc16 · inbound

Dismantling Pathological Shortcuts: A Causal Framework for Faithful LVLM Decoding cites this paper.

Dismantling Pathological Shortcuts: A Causal Framework for Faithful LVLM Decoding Object Hallucination in Image Captioning

Reference 77

Resolution
metadata mismatch
local_arxiv, observed 2026-07-01T19:06:02.557971Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-06-29T01:18:13.657975Z digest=sha256:66a3b927ef39bb78eb29b2e2bb8fd9dc14205f9845b59396f79187c71bc7c1a0

Observation 259d4117-c0f9-4fda-88f8-bdf868c4f613 · inbound

HKVLM: Faithful Query--Region Binding for Frozen-Detector Visual Grounding cites this paper.

HKVLM: Faithful Query--Region Binding for Frozen-Detector Visual Grounding Object Hallucination in Image Captioning

Reference 2

Resolution
unresolved
no resolver link, observed 2026-07-12T11:14:52.015620Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T11:14:52.015620Z digest=sha256:e1481ab0433a6aa763dd7d3b4f7d9d963975108f6a235f4614af28442dbe19a2

Observation 739c3fa0-0e2e-48cc-8e4a-3b7cbc5ec0b4 · inbound

Contextualized Evaluation of Vision Language Models through Dynamic, Multi-turn Interactions cites this paper.

Contextualized Evaluation of Vision Language Models through Dynamic, Multi-turn Interactions Object Hallucination in Image Captioning

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-02T01:57:58.861159Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T01:57:58.861159Z digest=sha256:bef57058c33ba113f1c9d8120d267cda72239dcf771f960bcd79db2099377fdb

Observation 8af5a2a3-2cf2-4f9f-bd96-cbaae4c706b9 · inbound

MissingBench-Verified: Probing Vision-Language Models' Inability to Detect Missing Object Parts cites this paper.

MissingBench-Verified: Probing Vision-Language Models' Inability to Detect Missing Object Parts Object Hallucination in Image Captioning

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-01T14:44:02.483130Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T14:44:02.483130Z digest=sha256:a73a58e3415ff89fdb988dd96ea0559ea668470c8890a4a7d22019a37d6a5d13

Observation abe41923-8cab-4434-aaa0-91b75873553d · inbound

C-PTQ: Fisher-weighted Channel-wise Sensitivity for Post-training Quantization of MLLMs cites this paper.

C-PTQ: Fisher-weighted Channel-wise Sensitivity for Post-training Quantization of MLLMs Object Hallucination in Image Captioning

Reference 127

Resolution
unresolved
no resolver link, observed 2026-08-01T08:36:00.831745Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T08:36:00.831745Z digest=sha256:215f861e2b1b2f5b9b2b384ca7180f827711fcaea6052f2709244d416f733ceb

Observation aa46c5ec-2e36-4cc1-badb-8f99e1a0b110 · inbound

Visual Token Compression Enhances Robustness of MLLMs cites this paper.

Visual Token Compression Enhances Robustness of MLLMs Object Hallucination in Image Captioning

Reference 68

Resolution
unresolved
no resolver link, observed 2026-08-01T13:10:37.549654Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T13:10:37.549654Z digest=sha256:b72c89fea64efaacdb03d37fcd9c640de374963ea7096902cb0051b8ee97c0a8

Observation ed379f1e-006a-4fda-887b-faef1b2908b6 · inbound

Disentangling Semantic Attention from Structural Bias in the Attention Manifold cites this paper.

Disentangling Semantic Attention from Structural Bias in the Attention Manifold Object Hallucination in Image Captioning

Reference 25

Resolution
unresolved
no resolver link, observed 2026-07-31T23:19:55.028375Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T23:19:55.028375Z digest=sha256:6d84c72b56e12308ea6d2debfbcf6fc0b9d939e7385d42150f99da14309f39cd