Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-04T13:27:59.419144Z
Paper Citation Record · LEDGER
As of 5 August 2026, this Paper Citation Record lists 27 of 27 outbound references and 4 inbound Pith citation observations for arXiv:2510.00705.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-04T13:27:59.419144Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-05T06:32:48.257954+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-02T20:30:57.931134Z
A source-named dated measurement, never combined with another source.
Source: pith, observed 2026-06-29T11:53:23.711158Z
27 of 27 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation fd28181a-4b7a-425d-80e1-0194aefd3762 · outbound
Training-free Uncertainty Guidance for Complex Visual Tasks with MLLMs (2017) were sampled at 3 FPS, while videos from ActivityNet Captions Krishna et al
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 32918eff-f67b-4b0f-b8f4-dff201ae4846 · outbound
Training-free Uncertainty Guidance for Complex Visual Tasks with MLLMs Qwen2.5-VL Technical Report
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1977dc0e-43d8-426e-852f-1019f5773a81 · outbound
Training-free Uncertainty Guidance for Complex Visual Tasks with MLLMs Language Models (Mostly) Know What They Know
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 65eacf5c-4fdb-4d56-b248-02c3a025184e · outbound
Training-free Uncertainty Guidance for Complex Visual Tasks with MLLMs TAG: A Simple Yet Effective Temporal-Aware Approach for Zero-Shot Video Temporal Grounding
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bcfa624a-27df-4efc-9fae-659e00ae8aab · outbound
Training-free Uncertainty Guidance for Complex Visual Tasks with MLLMs LLaVA-OneVision: Easy Visual Task Transfer
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 968f33d1-246c-4337-bb3f-7db71a08302a · outbound
Training-free Uncertainty Guidance for Complex Visual Tasks with MLLMs TextCoT: Zoom In for Enhanced Multimodal Text-Rich Image Understanding
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b2f74308-0b8f-4cda-a923-fb5fea0b3e70 · outbound
Training-free Uncertainty Guidance for Complex Visual Tasks with MLLMs Gemini 1.5: Unlocking multimodal understanding across millions of tokens of context
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 229c6e97-d91c-4bcd-a2ad-e567be55af63 · outbound
Training-free Uncertainty Guidance for Complex Visual Tasks with MLLMs The information bottleneck method
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation edc14937-629f-47b1-b776-1df37a0fe805 · outbound
Training-free Uncertainty Guidance for Complex Visual Tasks with MLLMs VaLiD: Mitigating the Hallucination of Large Vision Language Models by Visual Layer Fusion Contrastive Decoding
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 83e97f27-3ede-4322-9404-30dabd8d60ff · outbound
Training-free Uncertainty Guidance for Complex Visual Tasks with MLLMs InternVideo: General Video Foundation Models via Generative and Discriminative Learning
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 88ca922f-510e-4504-9f08-5e73b4c37d34 · outbound
Training-free Uncertainty Guidance for Complex Visual Tasks with MLLMs A Survey on Video Temporal Grounding with Multimodal Large Language Model
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c768bec0-c8c8-4a76-8389-60235b0a8898 · outbound
Training-free Uncertainty Guidance for Complex Visual Tasks with MLLMs Generate, but verify: Reducing hallucination in vision-language models with retrospective resam- pling.arXiv preprint arXiv:2504.13169, 2025b
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 63e2b82f-acc5-4c00-8289-ba68f89b3760 · outbound
Training-free Uncertainty Guidance for Complex Visual Tasks with MLLMs LMMs-Eval: Reality Check on the Evaluation of Large Multimodal Models
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4438ba32-5b43-4ea1-948c-2d2e64bf8975 · outbound
Training-free Uncertainty Guidance for Complex Visual Tasks with MLLMs DeepEyes: Incentivizing "Thinking with Images" via Reinforcement Learning
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e29c6fd0-4f3e-4ee1-848f-d899feec41dc · outbound
Training-free Uncertainty Guidance for Complex Visual Tasks with MLLMs From Seconds to Hours: Reviewing MultiModal Large Language Models on Comprehensive Long Video Understanding
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 77e04a20-0b5d-49c9-a544-ab6deaa77d44 · outbound
Training-free Uncertainty Guidance for Complex Visual Tasks with MLLMs Entropy, a foundational concept from information theory (Shannon, 1948), provides a formal measure of the uncertainty inherent in a probability distribution
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 824337d7-b091-4191-89f2-fd29d73d9074 · outbound
Training-free Uncertainty Guidance for Complex Visual Tasks with MLLMs Pretrained on massive text corpora, LLMs learn to generate reliable probability distributions over a predefined vocabulary
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 787f4816-e869-4bc5-9ded-b5ff21e63174 · outbound
Training-free Uncertainty Guidance for Complex Visual Tasks with MLLMs A” and “B
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e2eab76a-570b-4bac-a351-c243020a3693 · outbound
Training-free Uncertainty Guidance for Complex Visual Tasks with MLLMs LLaMA: Open and Efficient Foundation Language Models
Reference 2000
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 64e6e1a3-0bf5-4063-9b05-b25003ebc6f3 · outbound
Training-free Uncertainty Guidance for Complex Visual Tasks with MLLMs This finding aligns with principles from curriculum learning, where task difficulty can be measured by model uncertainty (Bengio et al., 2009; Kumar et al., 2010)
Reference 2008
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 09861003-066e-4122-a817-4db240f692bc · outbound
Training-free Uncertainty Guidance for Complex Visual Tasks with MLLMs The Internal State of an LLM Knows When It's Lying
Reference 2017
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 76dbb3a0-84e8-434c-9c6f-48b3029920af · outbound
Training-free Uncertainty Guidance for Complex Visual Tasks with MLLMs Threading Keyframe with Narratives: MLLMs as Strong Long Video Comprehenders
Reference 2019
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1f2463c6-462a-448c-822a-86e27085b6c5 · outbound
Training-free Uncertainty Guidance for Complex Visual Tasks with MLLMs TimeMarker: A Versatile Video-LLM for Long and Short Video Understanding with Superior Temporal Localization Ability
Reference 2020
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation af50635d-a1d9-4977-9180-3439cbf037d9 · outbound
Training-free Uncertainty Guidance for Complex Visual Tasks with MLLMs InternVideo2.5: Empowering Video MLLMs with Long and Rich Context Modeling
Reference 2022
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0e247db2-211e-49d4-8e58-e7bf5f68018e · outbound
Training-free Uncertainty Guidance for Complex Visual Tasks with MLLMs Qwen-VL: A Versatile Vision-Language Model for Understanding, Localization, Text Reading, and Beyond
Reference 2023
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b25ee4a3-3132-4bde-b546-d701ac7253df · outbound
Training-free Uncertainty Guidance for Complex Visual Tasks with MLLMs FRAG: Frame Selection Augmented Generation for Long Video and Long Document Understanding
Reference 2024
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a08925e4-ee69-435f-8e78-a70652459d75 · outbound
Training-free Uncertainty Guidance for Complex Visual Tasks with MLLMs GPT-4o System Card
Reference 2025
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 658fb90c-48e9-46b5-9ec5-7ac6969bc808 · inbound
LookWise: Knowing When and Where to Look for Fine-Grained Visual Reasoning in Multimodal Large Language Models Training-free Uncertainty Guidance for Complex Visual Tasks with MLLMs
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 98a8be6d-bab6-4ddb-a9ee-16a8dd9656c2 · inbound
Zoom Consistency: A Free Confidence Signal in Multi-Step Visual Grounding Pipelines Training-free Uncertainty Guidance for Complex Visual Tasks with MLLMs
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation d0ec393c-dffe-4757-af64-54b9d339b4d5 · inbound
DenoiseRL: Bootstrapping Reasoning Models to Recover from Noisy Prefixes Training-free Uncertainty Guidance for Complex Visual Tasks with MLLMs
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 3267df31-8efa-4278-a340-0a8425ae41b9 · inbound
DenoiseRL: Bootstrapping Reasoning Models to Recover from Noisy Prefixes Training-free Uncertainty Guidance for Complex Visual Tasks with MLLMs
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.