← back to paper
arxiv: 2506.17837 · 2 revisions
Time-Contrastive Pretraining for In-Context Image and Video Segmentation