Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 7 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 15 inbound Pith citation observations for arXiv:2407.14500.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-07T14:27:19.300329Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-07-04T15:09:55.216469Z
0 of 0 outbound references displayed
External citation measurements
No source-named external measurement is stored.
No outbound reference observations are available for this paper version.
Observation 59d82159-fd8b-4c8e-b1ef-e006c7f06415 · inbound
Reasoning Segmentation for Images and Videos: A Survey ViLLa: Video Reasoning Segmentation with Large Language Model
Reference 2024
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 51c3e5c4-7511-4120-a3b2-65b765c8c9b2 · inbound
InterRVOS: Interaction-aware Referring Video Object Segmentation ViLLa: Video Reasoning Segmentation with Large Language Model
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ab1af148-e5f5-45e1-86d7-1042694e9af6 · inbound
A Comprehensive Survey on Video Scene Parsing:Advances, Challenges, and Prospects ViLLa: Video Reasoning Segmentation with Large Language Model
Reference 200
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d69fd75a-9aa7-4bd6-928c-b428aaaf0e05 · inbound
DisCo: Towards Distinct and Coherent Visual Encapsulation in Video MLLMs ViLLa: Video Reasoning Segmentation with Large Language Model
Reference 88
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 51c74104-c7d0-4941-bc6f-b74ebc17b510 · inbound
HRSeg: High-Resolution Visual Perception and Enhancement for Reasoning Segmentation ViLLa: Video Reasoning Segmentation with Large Language Model
Reference 82
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f35ef5b4-7b2b-4406-9312-0d74b021a834 · inbound
Towards Omnimodal Expressions and Reasoning in Referring Audio-Visual Segmentation ViLLa: Video Reasoning Segmentation with Large Language Model
Reference 84
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation dd204a70-d312-42b5-b8a4-8b8f3464d726 · inbound
Unleashing Hierarchical Reasoning: An LLM-Driven Framework for Training-Free Referring Video Object Segmentation ViLLa: Video Reasoning Segmentation with Large Language Model
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b829bba1-a116-4cf6-a6c0-05e336f5e66c · inbound
LMMs Meet Object-Centric Vision: Understanding, Segmentation, Editing and Generation ViLLa: Video Reasoning Segmentation with Large Language Model
Reference 239
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 9752265a-1874-4753-a54e-5512d78a2de7 · inbound
PixDLM: A Dual-Path Multimodal Language Model for UAV Reasoning Segmentation ViLLa: Video Reasoning Segmentation with Large Language Model
Reference 61
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 41eec131-2845-444a-84fa-8165ba0eec00 · inbound
PixDLM: A Dual-Path Multimodal Language Model for UAV Reasoning Segmentation ViLLa: Video Reasoning Segmentation with Large Language Model
Reference 61
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 24c37d64-4d13-4d55-bf73-140cbe1b3500 · inbound
Weakly-Supervised Referring Video Object Segmentation through Text Supervision ViLLa: Video Reasoning Segmentation with Large Language Model
Reference 52
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 88f36ba8-30bd-4fa6-9265-6d15d5212724 · inbound
APRVOS: 1st Place Winner of 5th PVUW MeViS-Audio Track ViLLa: Video Reasoning Segmentation with Large Language Model
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 3b99a5eb-097e-40f2-838c-751e1f2a28ae · inbound
AgentRVOS for MeViS-Text Track of 5th PVUW Challenge: 3rd Method ViLLa: Video Reasoning Segmentation with Large Language Model
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 6e2118be-67cd-429d-b53a-d9e35271a891 · inbound
RCoT-Seg: Reinforced Chain-of-Thought for Video Reasoning and Segmentation ViLLa: Video Reasoning Segmentation with Large Language Model
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 7c45d39c-2791-40cc-8f92-915504edf8e5 · inbound
From Structure to Synergy: A Survey of Vision-Language Perception Paradigm Evolution in Multimodal Large Language Models ViLLa: Video Reasoning Segmentation with Large Language Model
Reference 84
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.