Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-05-10T12:32:11.911775Z
Paper Citation Record · LEDGER
As of 5 August 2026, this Paper Citation Record lists 29 of 29 outbound references and 1 inbound Pith citation observation for arXiv:2604.14630.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-05-10T12:32:11.911775Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-05T06:32:48.257954+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-05-10T12:32:11.911775Z
A source-named dated measurement, never combined with another source.
Source: pith, observed 2026-05-11T11:51:01.047996Z
29 of 29 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation cdf7d36e-0706-46f8-9420-a768af89e465 · outbound
CMTM: Cross-Modal Token Modulation for Unsupervised Video Object Segmentation CMTM: Cross-Modal Token Modulation for Unsupervised Video Object Segmentation
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation ad51c3f4-aace-410f-b90e-099fe6d021f3 · outbound
CMTM: Cross-Modal Token Modulation for Unsupervised Video Object Segmentation Two- stream architectures that combine these cues are widely ex- plored
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 3c0733f1-60e9-41ea-9b8c-2bcf0af0171d · outbound
CMTM: Cross-Modal Token Modulation for Unsupervised Video Object Segmentation Task Formulation In UVOS, the objective is to generate binary segmentation masksMfrom each input video sequence
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation c6cadc4d-14ec-4d2e-8bfb-2afefc7d5d34 · outbound
CMTM: Cross-Modal Token Modulation for Unsupervised Video Object Segmentation The evaluation datasets include the DA VIS 2016 [19] validation set (D), the FBMS [22] test set (F), the YouTube-Objects [23] (Y), and Long-Videos [24] dataset (L)
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation aebe1be2-a29c-4fd2-8573-095837f47e91 · outbound
CMTM: Cross-Modal Token Modulation for Unsupervised Video Object Segmentation CMTM outperforms state-of-the-art methods, demonstrating significant improvements in segmentation accuracy
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 92c40142-f587-48cb-9bce-431d360437ec · outbound
CMTM: Cross-Modal Token Modulation for Unsupervised Video Object Segmentation Matnet: Motion-attentive transition net- work for zero-shot video object segmentation
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation bc837542-53d2-4b28-9a53-8bf9c7e67572 · outbound
CMTM: Cross-Modal Token Modulation for Unsupervised Video Object Segmentation Full-duplex strategy for video object segmentation
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation b7e0050f-ba6e-4862-91e2-c149b924263a · outbound
CMTM: Cross-Modal Token Modulation for Unsupervised Video Object Segmentation Learning motion- appearance co-attention for zero-shot video object seg- mentation
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 87df2317-d50e-4335-9a6c-a7746ed8930f · outbound
CMTM: Cross-Modal Token Modulation for Unsupervised Video Object Segmentation Deep transport network for unsuper- vised video object segmentation
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 46c5a147-0b44-479e-8d9b-7e0a86d66614 · outbound
CMTM: Cross-Modal Token Modulation for Unsupervised Video Object Segmentation Reciprocal trans- formations for unsupervised video object segmentation
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 33491325-c423-4a94-83b4-bdec68083626 · outbound
CMTM: Cross-Modal Token Modulation for Unsupervised Video Object Segmentation Hierarchical feature alignment network for unsupervised video object seg- mentation
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 71db6232-bea3-4010-92de-9259e6f1a31b · outbound
CMTM: Cross-Modal Token Modulation for Unsupervised Video Object Segmentation Guided slot at- tention for unsupervised video object segmentation
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 4dd41f10-d90f-4bc6-8db3-361227d5f7d0 · outbound
CMTM: Cross-Modal Token Modulation for Unsupervised Video Object Segmentation Improving Unsupervised Video Object Segmentation via Fake Flow Generation
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 81693d5b-aa09-46a7-b4f4-4d8712b2dcc6 · outbound
CMTM: Cross-Modal Token Modulation for Unsupervised Video Object Segmentation Deep residual learning for image recognition
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation b7435767-ecce-442b-803e-cfd5eb8cb291 · outbound
CMTM: Cross-Modal Token Modulation for Unsupervised Video Object Segmentation D2conv3d: Dynamic dilated convolu- tions for object segmentation in videos
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 713050ee-5d23-41c0-86fc-9137755ad231 · outbound
CMTM: Cross-Modal Token Modulation for Unsupervised Video Object Segmentation Video classification with channel-separated con- volutional networks
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 7343eac6-c990-42f5-b0d0-662aaecf2e39 · outbound
CMTM: Cross-Modal Token Modulation for Unsupervised Video Object Segmentation Itera- tively selecting an easy reference frame makes unsuper- vised video object segmentation easier
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 8d674490-d77d-47b6-83a6-691fab5a1a13 · outbound
CMTM: Cross-Modal Token Modulation for Unsupervised Video Object Segmentation Segformer: Sim- ple and efficient design for semantic segmentation with transformers
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 150fcedc-7b4c-4b70-a1df-794470d256a8 · outbound
CMTM: Cross-Modal Token Modulation for Unsupervised Video Object Segmentation Unsupervised video object segmentation with online adversarial self-tuning
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation c4882ae3-b47b-4cb9-a9b0-f93aafcba563 · outbound
CMTM: Cross-Modal Token Modulation for Unsupervised Video Object Segmentation MobileViT: Light-weight, General-purpose, and Mobile-friendly Vision Transformer
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation c8b3e3a2-83eb-4270-b04a-9bf30f54772a · outbound
CMTM: Cross-Modal Token Modulation for Unsupervised Video Object Segmentation Simulflow: Simultaneously extract- ing feature and identifying target for unsupervised video object segmentation
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 2a0ff1cf-742d-4a52-99f2-a2fe87267039 · outbound
CMTM: Cross-Modal Token Modulation for Unsupervised Video Object Segmentation Generalizable fourier augmen- tation for unsupervised video object segmentation
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 1b41aa70-6c34-4efd-b843-4fb0c7de53ad · outbound
CMTM: Cross-Modal Token Modulation for Unsupervised Video Object Segmentation YouTube-VOS: A Large-Scale Video Object Segmentation Benchmark
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 7a0e0aec-6184-48f0-9c66-abb7c483ef82 · outbound
CMTM: Cross-Modal Token Modulation for Unsupervised Video Object Segmentation The 2017 DAVIS Challenge on Video Object Segmentation
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 10cf0872-39b5-4957-ac46-96b3bb7fa40c · outbound
CMTM: Cross-Modal Token Modulation for Unsupervised Video Object Segmentation Learning to detect salient objects with image-level supervision
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation cf4066db-e7be-4da9-a655-c8eb8b4932d4 · outbound
CMTM: Cross-Modal Token Modulation for Unsupervised Video Object Segmentation Adam: A Method for Stochastic Optimization
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 4572ec8e-8889-421d-aab7-13d0f6641db3 · outbound
CMTM: Cross-Modal Token Modulation for Unsupervised Video Object Segmentation Segmen- tation of moving objects by long term video analysis
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 3410c907-bb07-473a-b76c-d3381b030421 · outbound
CMTM: Cross-Modal Token Modulation for Unsupervised Video Object Segmentation Learning object class detectors from weakly annotated video
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 1617ad17-8a94-46cd-b153-ab8c60e88f70 · outbound
CMTM: Cross-Modal Token Modulation for Unsupervised Video Object Segmentation Video object segmentation with adaptive feature bank and uncertain-region refinement
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation cdf7d36e-0706-46f8-9420-a768af89e465 · inbound
CMTM: Cross-Modal Token Modulation for Unsupervised Video Object Segmentation CMTM: Cross-Modal Token Modulation for Unsupervised Video Object Segmentation
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.