Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-03T06:36:26.484082Z
Paper Citation Record · LEDGER
As of 18 August 2026, this Paper Citation Record lists 22 of 22 outbound references and 1 inbound Pith citation observation for arXiv:2601.22574.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-03T06:36:26.484082Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-18T06:34:40.430872+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-05-12T02:31:40.463891Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-05-12T07:36:31.820599Z
22 of 22 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 9eef630c-9d12-4349-937f-3ceb8eba3994 · outbound
Enhancing Video Representations with Spatiotemporal-Semantic Residual to Mitigate Hallucinations in Video Large Multimodal Models Grounding language with vision: A conditional mutual information calibrated decoding strat- egy for reducing hallucinations in lvlms.arXiv preprint arXiv:2505.19678,
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 47c1b3e6-57a4-4dab-94d9-c39b6b685622 · outbound
Enhancing Video Representations with Spatiotemporal-Semantic Residual to Mitigate Hallucinations in Video Large Multimodal Models Exploring Hallucination of Large Multimodal Models in Video Understanding: Benchmark, Analysis and Mitigation
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f68a30ba-060e-45b2-8058-16df557cbff3 · outbound
Enhancing Video Representations with Spatiotemporal-Semantic Residual to Mitigate Hallucinations in Video Large Multimodal Models Mentalmac: Enhancing large language models for detect- ing mental manipulation via multi-task anti-curriculum distillation.arXiv preprint arXiv:2505.15255, 2025b
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 41f780d3-8f56-4fd1-b112-2ccb19bdf0e7 · outbound
Enhancing Video Representations with Spatiotemporal-Semantic Residual to Mitigate Hallucinations in Video Large Multimodal Models VistaDPO: Video Hierarchical Spatial-Temporal Direct Preference Optimization for Large Video Models
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bd967670-d41d-46f1-b8f7-289c78e505c0 · outbound
Enhancing Video Representations with Spatiotemporal-Semantic Residual to Mitigate Hallucinations in Video Large Multimodal Models H., Jo, Y ., and Seo, M
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d5808543-5658-4a0d-a05f-415c4feb81f9 · outbound
Enhancing Video Representations with Spatiotemporal-Semantic Residual to Mitigate Hallucinations in Video Large Multimodal Models Decoupled Weight Decay Regularization
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation da3e0954-924c-4902-80ea-91d18c307674 · outbound
Enhancing Video Representations with Spatiotemporal-Semantic Residual to Mitigate Hallucinations in Video Large Multimodal Models Countervid: Counterfactual video generation for mitigat- ing action and temporal hallucinations in video-language models.arXiv preprint arXiv:2601.04778,
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4de1f7b8-f60c-4ef0-a021-b1bbfe3febcc · outbound
Enhancing Video Representations with Spatiotemporal-Semantic Residual to Mitigate Hallucinations in Video Large Multimodal Models Smart- sight: Mitigating hallucination in video-llms without com- promising video understanding via temporal attention collapse.arXiv preprint arXiv:2512.18671,
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d80fcd96-959f-4a72-9985-1e19fcb2a6e0 · outbound
Enhancing Video Representations with Spatiotemporal-Semantic Residual to Mitigate Hallucinations in Video Large Multimodal Models VideoHallucer: Evaluating Intrinsic and Extrinsic Hallucinations in Large Video-Language Models
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4183cc5b-88aa-481a-af8f-b4baa078a506 · outbound
Enhancing Video Representations with Spatiotemporal-Semantic Residual to Mitigate Hallucinations in Video Large Multimodal Models Unresolved cited work
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9498603f-c74d-4438-8ebf-1b6bc3972838 · outbound
Enhancing Video Representations with Spatiotemporal-Semantic Residual to Mitigate Hallucinations in Video Large Multimodal Models Kardia-r1: Unleashing llms to reason to- ward understanding and empathy for emotional support via rubric-as-judge reinforcement learning.arXiv preprint arXiv:2512.01282, 2025a
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d5c1c029-6c75-436f-b0b4-16d7c869f433 · outbound
Enhancing Video Representations with Spatiotemporal-Semantic Residual to Mitigate Hallucinations in Video Large Multimodal Models Video-llama: An instruction- tuned audio-visual language model for video understand- ing
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 96dd6031-5c4b-4cab-a7c0-83252aa71abb · outbound
Enhancing Video Representations with Spatiotemporal-Semantic Residual to Mitigate Hallucinations in Video Large Multimodal Models Eventhallusion: Diagnosing event hallucinations in video llms.arXiv preprint arXiv:2409.16597, 2024a
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 237834a3-6d1f-41ef-9563-5d0b483f3496 · outbound
Enhancing Video Representations with Spatiotemporal-Semantic Residual to Mitigate Hallucinations in Video Large Multimodal Models j., Gui, L., Fu, D., Feng, J., Liu, Z., and Li, C
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 659eaaf6-511a-473d-9f5c-9b76a50e4198 · outbound
Enhancing Video Representations with Spatiotemporal-Semantic Residual to Mitigate Hallucinations in Video Large Multimodal Models Can pruning improve reasoning? revisiting long-cot compression with capability in mind for better reasoning
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5e077945-d7b4-47c6-84b1-e83c05dd89c0 · outbound
Enhancing Video Representations with Spatiotemporal-Semantic Residual to Mitigate Hallucinations in Video Large Multimodal Models Layernorm Linear GeLU Linear GeLU Linear Tanh Layernorm Linear GeLU Linear GeLU Linear Tanh Figure 6.Overview of the architecture of SSD
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b7b21c44-bcda-46b7-982c-d67aab7717c5 · outbound
Enhancing Video Representations with Spatiotemporal-Semantic Residual to Mitigate Hallucinations in Video Large Multimodal Models Affordance-R1: Reinforcement Learning for Generalizable Affordance Reasoning in Multimodal Large Language Model
Reference 2017
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0a6763e4-2535-425b-9843-2594ba84dcf7 · outbound
Enhancing Video Representations with Spatiotemporal-Semantic Residual to Mitigate Hallucinations in Video Large Multimodal Models Helpd: Mitigating hallucination of lvlms by hierarchical feedback learning with vision-enhanced penalty decoding
Reference 2019
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 79c897ce-3b31-45da-9dc2-7ccd05ab297e · outbound
Enhancing Video Representations with Spatiotemporal-Semantic Residual to Mitigate Hallucinations in Video Large Multimodal Models Unresolved cited work
Reference 2022
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ddf08541-6493-4209-b923-2632e50ca56d · outbound
Enhancing Video Representations with Spatiotemporal-Semantic Residual to Mitigate Hallucinations in Video Large Multimodal Models Mitigating Hallucination in VideoLLMs via Temporal-Aware Activation Engineering
Reference 2023
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8f265e5e-95f3-4a29-a098-211c1d0d2a1c · outbound
Enhancing Video Representations with Spatiotemporal-Semantic Residual to Mitigate Hallucinations in Video Large Multimodal Models Decoupling contrastive decoding: Robust hallucination mitigation in multimodal large language models.arXiv preprint arXiv:2504.08809,
Reference 2024
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3c5176a0-b158-44a3-b733-7f69cc794a5b · outbound
Enhancing Video Representations with Spatiotemporal-Semantic Residual to Mitigate Hallucinations in Video Large Multimodal Models PaMi-VDPO: Mitigating Video Hallucinations by Prompt-Aware Multi-Instance Video Preference Learning
Reference 2025
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a41e6d9f-3019-400d-b960-51b86cea64b0 · inbound
Tracking the Truth: Object-Centric Spatio-Temporal Monitoring for Video Large Language Models Enhancing Video Representations with Spatiotemporal-Semantic Residual to Mitigate Hallucinations in Video Large Multimodal Models
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.