Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-08T05:58:12.511658Z
Paper Citation Record · LEDGER
As of 9 August 2026, this Paper Citation Record lists 13 of 13 outbound references and 0 inbound Pith citation observations for arXiv:2608.05592.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-08T05:58:12.511658Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
13 of 13 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 6def475b-324d-4410-88f9-f47b87a9bc99 · outbound
Beyond Frame Selection: Rethinking Long-Video Understanding with MLLMs PaLM 2 Technical Report
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ecaf5288-9df8-4e82-a509-3ba87408f1fc · outbound
Beyond Frame Selection: Rethinking Long-Video Understanding with MLLMs Gemini 1.5: Unlocking multimodal understanding across millions of tokens of context
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 62a6bca5-905c-4c51-a7e7-0ad0095cd7bc · outbound
Beyond Frame Selection: Rethinking Long-Video Understanding with MLLMs LLaVA-OneVision: Easy Visual Task Transfer
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0c75d6b4-4d1b-4ecc-80f8-c31288db3ffa · outbound
Beyond Frame Selection: Rethinking Long-Video Understanding with MLLMs Commonsense video question answering through video-grounded entailment tree reasoning
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2963d8cb-b5c7-4071-a8bd-c57a84f4359e · outbound
Beyond Frame Selection: Rethinking Long-Video Understanding with MLLMs Video-MTR: Reinforced Multi-Turn Reasoning for Long Video Understanding
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ee3767d9-c33d-43d2-bb66-3cda6ad7a77e · outbound
Beyond Frame Selection: Rethinking Long-Video Understanding with MLLMs Qwen2 Technical Report
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f23307aa-93c6-45e4-91f0-af19404e5a18 · outbound
Beyond Frame Selection: Rethinking Long-Video Understanding with MLLMs mPLUG-Owl: Modularization Empowers Large Language Models with Multimodality
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 22419d38-e134-4199-bea9-b3855609dab1 · outbound
Beyond Frame Selection: Rethinking Long-Video Understanding with MLLMs Thinking With Videos: Multimodal Tool-Augmented Reinforcement Learning for Long Video Reasoning
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 53c686ca-3b8d-43fd-be20-a2701b80165b · outbound
Beyond Frame Selection: Rethinking Long-Video Understanding with MLLMs Videolucy: Deep memory backtracking for long video understanding.arXiv:2510.12422,
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation dbf2bbd4-9c1f-40b3-8550-3b437f0bf446 · outbound
Beyond Frame Selection: Rethinking Long-Video Understanding with MLLMs 14 C.2 Ablation on the Global Representation Depth
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation b0d01f99-df2b-4f19-80e9-377357e4ccb3 · outbound
Beyond Frame Selection: Rethinking Long-Video Understanding with MLLMs Temporal Chain of Thought: Long-Video Understanding by Thinking in Frames
Reference 2023
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fffa4d07-097b-4e3e-aa76-2dfb45a3bd33 · outbound
Beyond Frame Selection: Rethinking Long-Video Understanding with MLLMs GPT-4o System Card
Reference 2024
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 884443e7-ef25-4171-a095-1834aa17c2ea · outbound
Beyond Frame Selection: Rethinking Long-Video Understanding with MLLMs MiniGPT4-Video: Advancing Multimodal LLMs for Video Understanding with Interleaved Visual-Textual Tokens
Reference 2025
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
No inbound Pith citation observations are available.