Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T11:29:29.724751Z
Paper Citation Record · LEDGER
As of 8 August 2026, this Paper Citation Record lists 32 of 32 outbound references and 1 inbound Pith citation observation for arXiv:2506.02356.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T11:29:29.724751Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-06-26T01:50:54.242508Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-07-04T15:09:55.213452Z
32 of 32 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation dee763a0-162e-4fdf-803b-272cdbdf7042 · outbound
InterRVOS: Interaction-aware Referring Video Object Segmentation One token to seg them all: Language instructed reasoning segmentation in videos
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 85184742-8ac3-4641-a574-aecbcd50b8c5 · outbound
InterRVOS: Interaction-aware Referring Video Object Segmentation End-to-end referring video object segmentation with multimodal transformers
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 5a381820-99cb-45bc-9192-783b00a298f8 · outbound
InterRVOS: Interaction-aware Referring Video Object Segmentation Expanding Performance Boundaries of Open-Source Multimodal Models with Model, Data, and Test-Time Scaling
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9368ec74-88b7-4f2f-a2c8-0c690f01fbda · outbound
InterRVOS: Interaction-aware Referring Video Object Segmentation Vision-language transformer and query generation for referring segmentation
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation baae64bd-df07-4a70-b212-5abaec036c7b · outbound
InterRVOS: Interaction-aware Referring Video Object Segmentation Mevis: A large-scale benchmark for video segmentation with motion expressions
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation b91dbc1a-cbb9-4d1f-aa8b-a8e0ac7e1dd8 · outbound
InterRVOS: Interaction-aware Referring Video Object Segmentation Moma: A multi-object multi-action dataset for understanding human activities
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation bf4a6e33-31c5-4bfb-ad76-b6f4018096bb · outbound
InterRVOS: Interaction-aware Referring Video Object Segmentation Actor and action video segmentation from a sentence
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 52848177-7ddd-4880-9b09-00a850dcaec7 · outbound
InterRVOS: Interaction-aware Referring Video Object Segmentation The llama 3 herd of models
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation f51f17a5-0d00-409d-b20a-94e1e8dfbb97 · outbound
InterRVOS: Interaction-aware Referring Video Object Segmentation Lora: Low-rank adaptation of large language models
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 82a39cc5-de55-4213-afa2-c3f0483ebaad · outbound
InterRVOS: Interaction-aware Referring Video Object Segmentation GPT-4o System Card
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5ad1d67c-d91c-4f6d-a940-9f0cc036ade9 · outbound
InterRVOS: Interaction-aware Referring Video Object Segmentation Action genome: Actions as compositions of spatiotemporal scene graphs
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 86063e0b-4cee-4535-af7a-e12a2afdbae7 · outbound
InterRVOS: Interaction-aware Referring Video Object Segmentation Video object segmentation with language referring expressions
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 3c42c50f-28d6-4f2a-a0c1-5d2179bed05c · outbound
InterRVOS: Interaction-aware Referring Video Object Segmentation VideoGLaMM: A Large Multimodal Model for Pixel-Level Visual Grounding in Videos
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9dd12239-b4e5-43e3-95f1-eab0bf2524d0 · outbound
InterRVOS: Interaction-aware Referring Video Object Segmentation Visual instruction tuning
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b8e90b91-2bcd-4c5a-b288-d8694652df00 · outbound
InterRVOS: Interaction-aware Referring Video Object Segmentation Spectrum-guided multi-granularity referring video object segmentation
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 945be114-8625-4e57-9bbe-9fc57529e925 · outbound
InterRVOS: Interaction-aware Referring Video Object Segmentation Refer-youtube-vos: A dataset for video object segmentation with language referring expressions
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 80d7bdab-9f61-4f66-a496-b15b3a3355e5 · outbound
InterRVOS: Interaction-aware Referring Video Object Segmentation SAM 2: Segment Anything in Images and Videos
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation db77b5ba-1be2-4a7a-9791-269ea4c563db · outbound
InterRVOS: Interaction-aware Referring Video Object Segmentation Urvos: Unified referring video object segmentation network with a large-scale benchmark
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 2aa89ca2-753c-40bd-b9e7-699f0531a945 · outbound
InterRVOS: Interaction-aware Referring Video Object Segmentation Video relationship detection
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 71050332-79e4-42da-8df2-789af37553ba · outbound
InterRVOS: Interaction-aware Referring Video Object Segmentation Annotating objects and relations in user-generated videos
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation e30af10b-5310-4e50-a02a-6535d18e2ab7 · outbound
InterRVOS: Interaction-aware Referring Video Object Segmentation Yfcc100m: The new data in multimedia research
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 5cb9c047-76ec-408a-8c66-0f6c3378604a · outbound
InterRVOS: Interaction-aware Referring Video Object Segmentation SOC: Semantic-Assisted Object Cluster for Referring Video Object Segmentation
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 51c3e5c4-7511-4120-a3b2-65b765c8c9b2 · outbound
InterRVOS: Interaction-aware Referring Video Object Segmentation ViLLa: Video Reasoning Segmentation with Large Language Model
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e701675b-9556-4572-b972-89f1ac135800 · outbound
InterRVOS: Interaction-aware Referring Video Object Segmentation Language as queries for referring video object segmentation
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 759cfd64-7b93-4d6e-9e2e-1b9fa7556250 · outbound
InterRVOS: Interaction-aware Referring Video Object Segmentation STAR: A Benchmark for Situated Reasoning in Real-World Videos
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3efc5d46-5ba5-4ef7-8314-11c0529b260c · outbound
InterRVOS: Interaction-aware Referring Video Object Segmentation Visa: Reasoning video object segmentation via large language models
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 651c05f3-96ee-4b49-9b2e-fae4fe859f67 · outbound
InterRVOS: Interaction-aware Referring Video Object Segmentation Sa2VA: Marrying SAM2 with LLaVA for Dense Grounded Understanding of Images and Videos
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f2776946-a890-4711-adbb-6403f95f64e9 · outbound
InterRVOS: Interaction-aware Referring Video Object Segmentation Decoupling Static and Hierarchical Motion Perception for Referring Video Segmentation
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 9034104e-4cb4-4125-b42f-18ec53fd65aa · outbound
InterRVOS: Interaction-aware Referring Video Object Segmentation write newline
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1d30c372-044f-4464-b690-037358a9bc37 · outbound
InterRVOS: Interaction-aware Referring Video Object Segmentation @esa (Ref
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c13ea854-4ba4-4081-8f51-25305ab4c2b2 · outbound
InterRVOS: Interaction-aware Referring Video Object Segmentation Unresolved cited work
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ce1d1cc0-77ce-4bad-b412-6a3faf66d064 · outbound
InterRVOS: Interaction-aware Referring Video Object Segmentation A child helping another child with a backpack
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 66f5d8ba-b7bd-4887-b54e-b3e839842af9 · inbound
From Structure to Synergy: A Survey of Vision-Language Perception Paradigm Evolution in Multimodal Large Language Models InterRVOS: Interaction-aware Referring Video Object Segmentation
Reference 90
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.