Pith. sign in

Paper Citation Record · LEDGER

VideoHallucer: Evaluating Intrinsic and Extrinsic Hallucinations in Large Video-Language Models

As of 7 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 21 inbound Pith citation observations for arXiv:2406.16338.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2406.16338 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 21 of 21 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 21 of 21 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T05:43:34.855411Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-03T10:27:56.151478Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 953d1368-6587-47f8-a72a-d3308781a7c1 · inbound

VidHal: Benchmarking Temporal Hallucinations in Vision LLMs cites this paper.

VidHal: Benchmarking Temporal Hallucinations in Vision LLMs VideoHallucer: Evaluating Intrinsic and Extrinsic Hallucinations in Large Video-Language Models

Reference 53

Resolution
verified exact
arxiv_id, observed 2026-05-23T16:58:12.093330Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-23T16:57:12.821916Z digest=sha256:cb761240fdef92eb46fdf922b4bff304ae912f8dcb3f1c2a7e2973fcacde4cd5

Observation 07f4a18e-c5b7-4a42-880f-b2a1811df3c7 · inbound

Mitigating Behavioral Hallucination in Multimodal Large Language Models for Sequential Images cites this paper.

Mitigating Behavioral Hallucination in Multimodal Large Language Models for Sequential Images VideoHallucer: Evaluating Intrinsic and Extrinsic Hallucinations in Large Video-Language Models

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-07T05:43:34.855411Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:43:34.855411Z digest=sha256:f47f006547ac308c949966cc4a49660368460964db9cd3470dab9b38ac7290fc

Observation d7d201a4-5e4f-4a6d-b11d-e3d890e46e7a · inbound

ARGUS: Hallucination and Omission Evaluation in Video-LLMs cites this paper.

ARGUS: Hallucination and Omission Evaluation in Video-LLMs VideoHallucer: Evaluating Intrinsic and Extrinsic Hallucinations in Large Video-Language Models

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-07T05:41:39.790898Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:41:39.790898Z digest=sha256:91afe6e4c7bd89a0e4c3d56c492cef578aea739313d639557fc62a512641d929

Observation ec42d65f-0f14-4190-a446-33c46f45a37f · inbound

FIFA: Unified Faithfulness Evaluation Framework for Text-to-Video and Video-to-Text Generation cites this paper.

FIFA: Unified Faithfulness Evaluation Framework for Text-to-Video and Video-to-Text Generation VideoHallucer: Evaluating Intrinsic and Extrinsic Hallucinations in Large Video-Language Models

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-06T19:06:15.872778Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T19:06:15.872778Z digest=sha256:34e138b65d1f40fc569da64ca2169ff838505baf1585f52f693c6befc2596e67

Observation 027c2733-a6de-4e19-b0b2-cc995b089554 · inbound

A comprehensive taxonomy of hallucinations in Large Language Models cites this paper.

A comprehensive taxonomy of hallucinations in Large Language Models VideoHallucer: Evaluating Intrinsic and Extrinsic Hallucinations in Large Video-Language Models

Reference 97

Resolution
unresolved
no resolver link, observed 2026-08-06T05:29:20.162727Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T05:29:20.162727Z digest=sha256:9b02f0a62d72c83c2eb2896ec42a164f8c684e1aab7887cbb945d1d28edc4a15

Observation 67395221-9853-4575-b4c2-3f03ae334068 · inbound

Empowering Multimodal LLMs with External Tools: A Comprehensive Survey cites this paper.

Empowering Multimodal LLMs with External Tools: A Comprehensive Survey VideoHallucer: Evaluating Intrinsic and Extrinsic Hallucinations in Large Video-Language Models

Reference 77

Resolution
unresolved
no resolver link, observed 2026-08-05T20:28:49.973347Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T20:28:49.973347Z digest=sha256:cbe508af34dc83d3c3e27951220490da55d004af0866aa8f9a0d3065bcd33174

Observation 14760259-ed39-4b68-b20a-974d045ae3cb · inbound

MESH -- Understanding Videos Like Human: Measuring Hallucinations in Large Video Models cites this paper.

MESH -- Understanding Videos Like Human: Measuring Hallucinations in Large Video Models VideoHallucer: Evaluating Intrinsic and Extrinsic Hallucinations in Large Video-Language Models

Reference 84

Resolution
unresolved
no resolver link, observed 2026-08-04T20:32:56.952259Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:32:56.952259Z digest=sha256:efda25a78f293428eaed406122051d4edd62a0fbb9a7b11e8bcb2a673b9def24

Observation d80fcd96-959f-4a72-9985-1e19fcb2a6e0 · inbound

Enhancing Video Representations with Spatiotemporal-Semantic Residual to Mitigate Hallucinations in Video Large Multimodal Models cites this paper.

Enhancing Video Representations with Spatiotemporal-Semantic Residual to Mitigate Hallucinations in Video Large Multimodal Models VideoHallucer: Evaluating Intrinsic and Extrinsic Hallucinations in Large Video-Language Models

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-03T06:36:26.055855Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T06:36:26.055855Z digest=sha256:6ed9d3e907232aa0f8eeb7aec5f0d6b7d27415de7c11d5f64b4210480f1720b1

Observation c60caf0e-8f36-4278-a306-886a66ec9cd3 · inbound

Relaxing Anchor-Frame Dominance for Mitigating Hallucinations in Video Large Language Models cites this paper.

Relaxing Anchor-Frame Dominance for Mitigating Hallucinations in Video Large Language Models VideoHallucer: Evaluating Intrinsic and Extrinsic Hallucinations in Large Video-Language Models

Reference 33

Resolution
verified exact
arxiv_id, observed 2026-05-11T11:26:03.552575Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-10T14:54:52.710686Z digest=sha256:0819d29ffff656e769e01d602e8a15bb8130194d87006a3c7923ba35e9c2a2f1

Observation dc1ac5f3-e5f9-4bb0-be03-28ec57215932 · inbound

When Text Hijacks Vision: Benchmarking and Mitigating Text Overlay-Induced Hallucination in Vision Language Models cites this paper.

When Text Hijacks Vision: Benchmarking and Mitigating Text Overlay-Induced Hallucination in Vision Language Models VideoHallucer: Evaluating Intrinsic and Extrinsic Hallucinations in Large Video-Language Models

Reference 62

Resolution
verified exact
arxiv_id, observed 2026-05-10T06:46:37.525334Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-10T06:41:59.641410Z digest=sha256:05bf82c8b86b28f39e09b9f4d9b31a59ac08ff88da722c39b218e597e7c79b09

Observation 233d9c0f-9cb1-4ffa-b4e5-83eee116964d · inbound

Raven: Rethinking Automated Assessment for Scratch Programs via Video-Grounded Evaluation cites this paper.

Raven: Rethinking Automated Assessment for Scratch Programs via Video-Grounded Evaluation VideoHallucer: Evaluating Intrinsic and Extrinsic Hallucinations in Large Video-Language Models

Reference 63

Resolution
verified exact
arxiv_id, observed 2026-05-10T11:10:09.917066Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-10T04:52:42.032751Z digest=sha256:f36fa6379e1fb58aad9d2484cea15d3b7e294b25a7cdf9e34757438ee882e660

Observation 8ba7f7f4-e5b6-4f15-8692-f2d1bc920c28 · inbound

Video-ToC: Video Tree-of-Cue Reasoning cites this paper.

Video-ToC: Video Tree-of-Cue Reasoning VideoHallucer: Evaluating Intrinsic and Extrinsic Hallucinations in Large Video-Language Models

Reference 15

Resolution
verified exact
arxiv_id, observed 2026-05-11T13:36:09.697431Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-10T01:20:29.374012Z digest=sha256:b206f177744f11206f2311221354978bced1f6524a6d434a758376d358355160

Observation 647867db-5ae3-4a31-b34e-049a3bf89270 · inbound

Towards Temporal Compositional Reasoning in Long-Form Sports Videos cites this paper.

Towards Temporal Compositional Reasoning in Long-Form Sports Videos VideoHallucer: Evaluating Intrinsic and Extrinsic Hallucinations in Large Video-Language Models

Reference 35

Resolution
unresolved
no resolver link, observed 2026-07-14T19:27:29.843866Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T19:27:29.843866Z digest=sha256:8cdf6b3325386cf95c094cb2dfdc6760e4918c6c7a23f4430d17c1a1799f27ae

Observation 534d4747-67b5-42ac-b4ed-b35df1bc28f7 · inbound

From Priors to Perception: Grounding Video-LLMs in Physical Reality cites this paper.

From Priors to Perception: Grounding Video-LLMs in Physical Reality VideoHallucer: Evaluating Intrinsic and Extrinsic Hallucinations in Large Video-Language Models

Reference 38

Resolution
verified exact
arxiv_id, observed 2026-05-11T17:21:08.287569Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-08T17:41:23.233366Z digest=sha256:6e89bf9d9ff0987e6c3d457de45f7b32248dd99abab096e45c05926462abd7eb

Observation e0c4e2a0-238a-45a5-9bf0-b0dd0ba8ad08 · inbound

TOC-Bench: A Temporal Object Consistency Benchmark for Video Large Language Models cites this paper.

TOC-Bench: A Temporal Object Consistency Benchmark for Video Large Language Models VideoHallucer: Evaluating Intrinsic and Extrinsic Hallucinations in Large Video-Language Models

Reference 40

Resolution
verified exact
arxiv_id, observed 2026-05-12T06:31:26.781676Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-12T04:13:21.487431Z digest=sha256:eef3f12116338bdec08aa0acfccf0811c26d45b7005e11528b56bdc2a19b661a

Observation d1592a4c-4973-4d8d-9ed4-49b108b48f40 · inbound

TOC-Bench: A Temporal Object Consistency Benchmark for Video Large Language Models cites this paper.

TOC-Bench: A Temporal Object Consistency Benchmark for Video Large Language Models VideoHallucer: Evaluating Intrinsic and Extrinsic Hallucinations in Large Video-Language Models

Reference 40

Resolution
verified exact
arxiv_id, observed 2026-05-13T06:57:28.252738Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-13T06:53:42.726350Z digest=sha256:37a22f1701470028467f4999e2e5318d7b178dcb1aec7011fff790d03c2cd52d

Observation d7e427cf-6df5-443c-b8ee-d355a7adc5d1 · inbound

When Vision Speaks for Sound cites this paper.

When Vision Speaks for Sound VideoHallucer: Evaluating Intrinsic and Extrinsic Hallucinations in Large Video-Language Models

Reference 64

Resolution
verified exact
arxiv_id, observed 2026-05-20T22:13:46.830651Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-20T22:12:52.160596Z digest=sha256:01145c9dac4c6cccb42c1115c605b8a7c3247602cb6c4f2d777a037f17a7c55e

Observation 800a68a7-01b2-4ef0-9d04-a977a9a6ee02 · inbound

OmniHalluc-L: Counterfactual Benchmarking and Modality-Perturbation Reliability Calibration for Long-Form Omni Hallucination cites this paper.

OmniHalluc-L: Counterfactual Benchmarking and Modality-Perturbation Reliability Calibration for Long-Form Omni Hallucination VideoHallucer: Evaluating Intrinsic and Extrinsic Hallucinations in Large Video-Language Models

Reference 2

Resolution
verified exact
arxiv_id, observed 2026-07-02T06:06:41.663618Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-28T07:31:51.780636Z digest=sha256:b5329d952e62574149976962c42a65aacc41346857ef889e2a934d338d1f7e47

Observation 45772a24-4efd-4ff5-8619-d861dc3a8784 · inbound

MultiToP: Learning to Patch Visual Tokens to Mitigate Hallucinations in Video Large Multimodal Models cites this paper.

MultiToP: Learning to Patch Visual Tokens to Mitigate Hallucinations in Video Large Multimodal Models VideoHallucer: Evaluating Intrinsic and Extrinsic Hallucinations in Large Video-Language Models

Reference 17

Resolution
metadata mismatch
arxiv_id, observed 2026-07-03T10:27:56.152859Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-06-27T10:02:58.341050Z digest=sha256:f93a42108b785aaeac700b303e4566192c9496fcebf1b5b5e92917fd3e3709b3

Observation 14e98e60-218d-4b8f-b609-c45d9003ed08 · inbound

No Place to Hide: Benchmarking Video Hallucination with Background-Controlled Pairs cites this paper.

No Place to Hide: Benchmarking Video Hallucination with Background-Controlled Pairs VideoHallucer: Evaluating Intrinsic and Extrinsic Hallucinations in Large Video-Language Models

Reference 71

Resolution
metadata mismatch
arxiv_id, observed 2026-07-01T10:25:41.367812Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-07-01T05:35:08.200219Z digest=sha256:3d20258a4b917df51412efa1bdbf3f8132a806c3d44eeb8a7c81fca743c7e19b

Observation 59a397f0-75c9-4156-bc67-378eb53aaa3b · inbound

MoHallBench: A Benchmark for Motion Hallucination in Video Large Language Models cites this paper.

MoHallBench: A Benchmark for Motion Hallucination in Video Large Language Models VideoHallucer: Evaluating Intrinsic and Extrinsic Hallucinations in Large Video-Language Models

Reference 38

Resolution
metadata mismatch
arxiv_id, observed 2026-07-02T13:46:58.702099Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-07-02T13:44:48.250838Z digest=sha256:5cc20a08b21f5cf9f0c2e3147c9633e9dbf3fad93dbbcda2b87d7663f0a99df8