Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-05-20T05:23:32.202297Z
Paper Citation Record · LEDGER
As of 16 August 2026, this Paper Citation Record lists 77 of 77 outbound references and 2 inbound Pith citation observations for arXiv:2605.20110.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-05-20T05:23:32.202297Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-16T06:30:59.297886+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-05T04:26:29.380459Z
A source-named dated measurement, never combined with another source.
Source: pith, observed 2026-07-04T00:19:13.707397Z
77 of 77 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 3eb87b22-ca74-4fd6-a80d-0625405867dc · outbound
SetCon: Towards Open-Ended Referring Segmentation via Set-Level Concept Prediction Qwen3-VL Technical Report
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 76a43aa4-fa76-4894-a94b-cc6ef11e4f70 · outbound
SetCon: Towards Open-Ended Referring Segmentation via Set-Level Concept Prediction One token to seg them all: Language instructed reasoning segmentation in videos.Advances in Neural Information Processing Systems, pages 6833–6859
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 537249cd-a9b6-4950-bb87-f14a52f8a6b8 · outbound
SetCon: Towards Open-Ended Referring Segmentation via Set-Level Concept Prediction Sam 3: Segment anything with concepts
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation b84410cd-2886-4581-b02f-e5dd1de15705 · outbound
SetCon: Towards Open-Ended Referring Segmentation via Set-Level Concept Prediction End-to-end object detection with transformers
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation b79a0389-ccdb-473c-8194-0d3e54b72a31 · outbound
SetCon: Towards Open-Ended Referring Segmentation via Set-Level Concept Prediction Sam4mllm: Enhance multi-modal large language model for referring expression segmentation
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation adb172a5-1d0f-476b-bde4-aa1fc1c2ed7e · outbound
SetCon: Towards Open-Ended Referring Segmentation via Set-Level Concept Prediction Samwise: Infusing wisdom in sam2 for text-driven video segmentation
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 4dd02138-6991-4721-a4f2-7517e19b8fc8 · outbound
SetCon: Towards Open-Ended Referring Segmentation via Set-Level Concept Prediction Mevis: A large-scale benchmark for video segmentation with motion expressions
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation f0f2a0f5-8269-4eda-a32c-4a73f2323b20 · outbound
SetCon: Towards Open-Ended Referring Segmentation via Set-Level Concept Prediction Mevis: A multi-modal dataset for referring motion expression video segmentation.IEEE Transactions on Pattern Analysis and Machine Intelligence
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 738c484c-458b-4d6f-8100-226a69cd7b6e · outbound
SetCon: Towards Open-Ended Referring Segmentation via Set-Level Concept Prediction Vlt: Vision-language transformer and query generation for referring segmentation.IEEE Transactions on Pattern Analysis and Machine Intelligence, pages 7900–7916
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation eba51d53-d74b-4f53-876f-594b053c060d · outbound
SetCon: Towards Open-Ended Referring Segmentation via Set-Level Concept Prediction Sam2long: Enhancing sam 2 for long video segmentation with a training-free memory tree
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation f8c2b1c0-d4c3-4e68-b209-7911ffe5e812 · outbound
SetCon: Towards Open-Ended Referring Segmentation via Set-Level Concept Prediction Language-bridged spatial-temporal interaction for referring video object segmentation
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 0030514d-a3ea-4d51-a3e7-bdf183d20c61 · outbound
SetCon: Towards Open-Ended Referring Segmentation via Set-Level Concept Prediction Palm-e: an embodied multimodal language model
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 65221503-1609-4bc9-af69-db1be65ff5b8 · outbound
SetCon: Towards Open-Ended Referring Segmentation via Set-Level Concept Prediction The devil is in temporal token: High quality video reasoning segmentation
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation d5632eb5-18fe-4e92-b789-a9b13b3c8bfd · outbound
SetCon: Towards Open-Ended Referring Segmentation via Set-Level Concept Prediction Anomalygpt: Detecting industrial anomalies using large vision-language models
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation b932d839-ba8d-44d2-99dc-3c6007306863 · outbound
SetCon: Towards Open-Ended Referring Segmentation via Set-Level Concept Prediction Html: Hybrid temporal-scale multimodal learning framework for referring video object segmentation
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 70349b3b-6e2d-485f-9644-c7d2ba1eebfc · outbound
SetCon: Towards Open-Ended Referring Segmentation via Set-Level Concept Prediction Decoupling static and hierarchical motion perception for referring video segmentation
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 8b2e50b9-1027-4c6c-a371-caf6020b12df · outbound
SetCon: Towards Open-Ended Referring Segmentation via Set-Level Concept Prediction Segmentation from natural language expressions
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation c817f34c-adf7-4e4d-9535-e79b01cffcd4 · outbound
SetCon: Towards Open-Ended Referring Segmentation via Set-Level Concept Prediction Beyond one-to-one: Rethinking the referring image segmentation
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 4f5d4a88-bb00-459d-90bc-c0eaaac131f5 · outbound
SetCon: Towards Open-Ended Referring Segmentation via Set-Level Concept Prediction Densely connected parameter-efficient tuning for referring image segmentation
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation f5788793-933b-49b6-ba1f-163b149463a7 · outbound
SetCon: Towards Open-Ended Referring Segmentation via Set-Level Concept Prediction Mmr: A large-scale benchmark dataset for multi-target and multi-granularity reasoning segmentation
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation f3999a03-7346-464d-966d-97fe73e50016 · outbound
SetCon: Towards Open-Ended Referring Segmentation via Set-Level Concept Prediction Mdetr-modulated detection for end-to-end multi-modal understanding
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 9efdeed2-871b-44f1-845b-42d3e1208133 · outbound
SetCon: Towards Open-Ended Referring Segmentation via Set-Level Concept Prediction Referitgame: Referring to objects in photographs of natural scenes
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 19624c22-4726-42c1-830d-1c948fc3fbb2 · outbound
SetCon: Towards Open-Ended Referring Segmentation via Set-Level Concept Prediction Video object segmentation with referring expressions
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 3936b7c1-8714-46ea-b373-06499cc60387 · outbound
SetCon: Towards Open-Ended Referring Segmentation via Set-Level Concept Prediction Segment anything
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation b9ad5cbe-4a8d-470e-8645-70df876210ea · outbound
SetCon: Towards Open-Ended Referring Segmentation via Set-Level Concept Prediction Lisa: Reasoning segmentation via large language model
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation fdc7504d-0d66-48c7-b07f-19182bec6635 · outbound
SetCon: Towards Open-Ended Referring Segmentation via Set-Level Concept Prediction Text4seg: Reimagining image segmentation as text generation
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 19bc26ee-9c9f-4c10-b491-539122e0fced · outbound
SetCon: Towards Open-Ended Referring Segmentation via Set-Level Concept Prediction Grounded language-image pre-training
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 81cef537-b0dd-4be5-840e-8d5eda63d0b2 · outbound
SetCon: Towards Open-Ended Referring Segmentation via Set-Level Concept Prediction Gligen: Open-set grounded text-to-image generation
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 60a05966-5391-4ded-bcda-7b5eb35ac078 · outbound
SetCon: Towards Open-Ended Referring Segmentation via Set-Level Concept Prediction Open-vocabulary semantic segmentation with mask-adapted clip
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation c6d23df5-ee9e-4003-b5e3-cc0a71c9982e · outbound
SetCon: Towards Open-Ended Referring Segmentation via Set-Level Concept Prediction Glus: Global-local reasoning unified into a single large language model for video segmentation
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation b6194e08-f90d-4b0a-abb2-a6ffd5d65145 · outbound
SetCon: Towards Open-Ended Referring Segmentation via Set-Level Concept Prediction Gres: Generalized referring expression segmentation
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation c83c8857-48a1-4776-8a01-57631030330c · outbound
SetCon: Towards Open-Ended Referring Segmentation via Set-Level Concept Prediction Recurrent multimodal interaction for referring image segmentation
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation dbca0f50-7a81-4537-b173-a6b8d49d175a · outbound
SetCon: Towards Open-Ended Referring Segmentation via Set-Level Concept Prediction Grounding dino: Marrying dino with grounded pre-training for open-set object detection
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation fe49c8e2-58e5-42b8-872d-80c810ca4bcc · outbound
SetCon: Towards Open-Ended Referring Segmentation via Set-Level Concept Prediction Universal segmen- tation at arbitrary granularity with language instruction
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 80303571-30e8-4450-9717-77803b3819d2 · outbound
SetCon: Towards Open-Ended Referring Segmentation via Set-Level Concept Prediction Seg-Zero: Reasoning-Chain Guided Segmentation via Cognitive Reinforcement
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 2dd39f48-4e75-447e-8bfd-b9b880995e5a · outbound
SetCon: Towards Open-Ended Referring Segmentation via Set-Level Concept Prediction Visionreasoner: Unified reasoning-integrated visual perception via reinforcement learning
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 834d53f6-83ad-4dba-a2d5-c104decc2f56 · outbound
SetCon: Towards Open-Ended Referring Segmentation via Set-Level Concept Prediction Rsvp: Reasoning segmentation via visual prompting and multi-modal chain-of-thought
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation a695338d-6c50-43dc-9a4c-ba3fb332c8df · outbound
SetCon: Towards Open-Ended Referring Segmentation via Set-Level Concept Prediction Image segmentation using text and image prompts
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 2c35af56-c239-48df-b971-b5e6e1a632ba · outbound
SetCon: Towards Open-Ended Referring Segmentation via Set-Level Concept Prediction Soc: Semantic-assisted object cluster for referring video object segmentation.Advances in Neural Information Processing Systems, pages 26425–26437
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation b1a6f9c8-4336-4365-b003-6975ea7f4923 · outbound
SetCon: Towards Open-Ended Referring Segmentation via Set-Level Concept Prediction Generation and comprehension of unambiguous object descriptions
Reference 40
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 672f137d-8739-46da-bae2-9a535781bb13 · outbound
SetCon: Towards Open-Ended Referring Segmentation via Set-Level Concept Prediction Spectrum-guided multi-granularity referring video object segmentation
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 1f66f7e8-060b-4b29-9689-8f660dfbf472 · outbound
SetCon: Towards Open-Ended Referring Segmentation via Set-Level Concept Prediction Simple open- vocabulary object detection
Reference 42
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation a67dd0cc-66af-40fc-b756-9ff5700963ab · outbound
SetCon: Towards Open-Ended Referring Segmentation via Set-Level Concept Prediction Videoglamm: A large multimodal model for pixel-level visual grounding in videos
Reference 43
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 3bc5c6b4-7b0c-4ab4-b1bc-54d25f6b324b · outbound
SetCon: Towards Open-Ended Referring Segmentation via Set-Level Concept Prediction Glamm: Pixel grounding large multimodal model
Reference 44
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation e236a7b2-2272-4002-940f-2fe5098c548b · outbound
SetCon: Towards Open-Ended Referring Segmentation via Set-Level Concept Prediction Sam 2: Segment anything in images and videos
Reference 45
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 0f35b2f6-cb35-4e32-83e4-828eb8f77faf · outbound
SetCon: Towards Open-Ended Referring Segmentation via Set-Level Concept Prediction Pixellm: Pixel reasoning with large multimodal model
Reference 46
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation bcd25cdd-85c4-4e87-b2c4-8fc166944435 · outbound
SetCon: Towards Open-Ended Referring Segmentation via Set-Level Concept Prediction Urvos: Unified referring video object segmentation network with a large-scale benchmark
Reference 47
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 234bfc0a-998d-41e6-87bc-17a797e80f6c · outbound
SetCon: Towards Open-Ended Referring Segmentation via Set-Level Concept Prediction OpenAI GPT-5 System Card
Reference 48
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 20161db5-25b2-4471-800d-308692c2420a · outbound
SetCon: Towards Open-Ended Referring Segmentation via Set-Level Concept Prediction Gemini 1.5: Unlocking multimodal understanding across millions of tokens of context
Reference 49
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 03943dfa-0224-4878-be9e-c5d863b2ab67 · outbound
SetCon: Towards Open-Ended Referring Segmentation via Set-Level Concept Prediction Videoanydoor: High-fidelity video object insertion with precise motion control
Reference 50
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation da2f5a62-fd98-47ae-b2f3-7f058730f80e · outbound
SetCon: Towards Open-Ended Referring Segmentation via Set-Level Concept Prediction X-sam: From segment anything to any segmentation
Reference 51
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 4d8528cd-c447-4f7c-ac86-a99839fa929f · outbound
SetCon: Towards Open-Ended Referring Segmentation via Set-Level Concept Prediction Unlocking the Potential of MLLMs in Referring Expression Segmentation via a Light-weight Mask Decoder
Reference 52
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 7db59666-b6f0-46df-847c-40edae6a3cbf · outbound
SetCon: Towards Open-Ended Referring Segmentation via Set-Level Concept Prediction Un- veiling parts beyond objects: Towards finer-granularity referring expression segmentation
Reference 53
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation d7bbf721-287f-4ca6-a965-1586a5e4ff77 · outbound
SetCon: Towards Open-Ended Referring Segmentation via Set-Level Concept Prediction Deforming videos to masks: Flow matching for referring video segmentation
Reference 54
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 7be87b18-f613-439b-83b8-ec5a536e582d · outbound
SetCon: Towards Open-Ended Referring Segmentation via Set-Level Concept Prediction Hyperseg: Hybrid segmentation assistant with fine-grained visual perceiver
Reference 55
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation a8279055-4d80-4130-8b67-e99c39476047 · outbound
SetCon: Towards Open-Ended Referring Segmentation via Set-Level Concept Prediction Instructseg: Unifying instructed visual segmentation with multi-modal large language models
Reference 56
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation acfc1fc8-4642-44fb-afd1-a42dacd1d863 · outbound
SetCon: Towards Open-Ended Referring Segmentation via Set-Level Concept Prediction Onlinerefer: A simple online baseline for referring video object segmentation
Reference 57
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation f4e5db51-586b-4f44-b5b5-598e4646e75b · outbound
SetCon: Towards Open-Ended Referring Segmentation via Set-Level Concept Prediction Language as queries for referring video object segmentation
Reference 58
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 116540d0-c105-4183-be42-9e954e2d362e · outbound
SetCon: Towards Open-Ended Referring Segmentation via Set-Level Concept Prediction General object foundation model for images and videos at scale
Reference 59
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation df6a6a94-570c-478d-8ca7-9a71c2f48965 · outbound
SetCon: Towards Open-Ended Referring Segmentation via Set-Level Concept Prediction Gsva: Generalized segmentation via multimodal large language models
Reference 60
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 76775ba7-cdcd-4bd6-bb84-777c3a63fb9a · outbound
SetCon: Towards Open-Ended Referring Segmentation via Set-Level Concept Prediction Region-based cluster discrimination for visual representation learning
Reference 61
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 62034e00-7601-49b4-bf7a-e811216db534 · outbound
SetCon: Towards Open-Ended Referring Segmentation via Set-Level Concept Prediction Viddar: Vision language model-based task-detrimental content detection for augmented reality.IEEE transactions on visualization and computer graphics
Reference 62
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 9546c80c-1e60-47b5-8233-748e50e06bb4 · outbound
SetCon: Towards Open-Ended Referring Segmentation via Set-Level Concept Prediction Visa: Reasoning video object segmentation via large language models
Reference 63
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 81062f97-ad17-4ed4-8ca4-92d2b11d0419 · outbound
SetCon: Towards Open-Ended Referring Segmentation via Set-Level Concept Prediction Lavt: Language- aware vision transformer for referring image segmentation
Reference 64
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 16ae8d62-a0a5-4cce-80f9-7d80c41808cc · outbound
SetCon: Towards Open-Ended Referring Segmentation via Set-Level Concept Prediction Mattnet: Modular attention network for referring expression comprehension
Reference 65
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 52b6c867-b539-4fce-9dd7-63f49708b868 · outbound
SetCon: Towards Open-Ended Referring Segmentation via Set-Level Concept Prediction Sa2VA: Marrying SAM2 with LLaVA for Dense Grounded Understanding of Images and Videos
Reference 66
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation fcc88c27-156d-444c-ae29-fa6c9f709434 · outbound
SetCon: Towards Open-Ended Referring Segmentation via Set-Level Concept Prediction Omg-llava: Bridging image-level, object-level, pixel-level reasoning and understanding
Reference 67
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation c36a9fb7-f75a-4e05-bf97-86dff3564869 · outbound
SetCon: Towards Open-Ended Referring Segmentation via Set-Level Concept Prediction EVF-SAM: Early Vision-Language Fusion for Text-Prompted Segment Anything Model
Reference 68
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 13afa7f0-fe46-441f-a54b-7d2a3ca7572f · outbound
SetCon: Towards Open-Ended Referring Segmentation via Set-Level Concept Prediction Psalm: Pixelwise segmentation with large multi-modal model
Reference 69
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 349c8681-0bb2-4139-93a2-95bbafd5d4c2 · outbound
SetCon: Towards Open-Ended Referring Segmentation via Set-Level Concept Prediction Sec: Advancing complex video object segmentation via progressive concept construction
Reference 70
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 4e0a5877-bd1d-43d9-bcf8-1291b17b8e2a · outbound
SetCon: Towards Open-Ended Referring Segmentation via Set-Level Concept Prediction Villa: Video reasoning segmentation with large language model
Reference 71
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation e058cb1e-2a8c-41ea-9554-3320c1480a85 · outbound
SetCon: Towards Open-Ended Referring Segmentation via Set-Level Concept Prediction Regionclip: Region-based language-image pretraining
Reference 72
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation d47a6581-3349-4eb0-85c8-2920f6ea51b3 · outbound
SetCon: Towards Open-Ended Referring Segmentation via Set-Level Concept Prediction Tracking with Human-Intent Reasoning
Reference 73
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 55fdb31c-232f-46e5-88c3-99b24d8d29c3 · outbound
SetCon: Towards Open-Ended Referring Segmentation via Set-Level Concept Prediction Training-free spatio-temporal decoupled reasoning video segmentation with adaptive object memory
Reference 74
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation f98126f4-41db-4573-9b9c-a300445a68e5 · outbound
SetCon: Towards Open-Ended Referring Segmentation via Set-Level Concept Prediction Rt-2: Vision-language-action models transfer web knowledge to robotic control
Reference 75
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 3c562250-8c72-4648-85d5-acefb3f00ea9 · outbound
SetCon: Towards Open-Ended Referring Segmentation via Set-Level Concept Prediction Generalized decoding for pixel, image, and language
Reference 76
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation e6957b40-145e-4170-95a5-3e24172a1587 · outbound
SetCon: Towards Open-Ended Referring Segmentation via Set-Level Concept Prediction Segment everything everywhere all at once.Advances in neural information processing systems, 36:19769–19782
Reference 77
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 0db5790d-6502-408a-aaf6-ed4cafb42c71 · inbound
Beyond the Current Observation: Evaluating Multimodal Large Language Models in Controllable Non-Markov Games SetCon: Towards Open-Ended Referring Segmentation via Set-Level Concept Prediction
Reference 94
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 0d7abf58-df95-434f-8ac0-5523cb524756 · inbound
Rethinking Self-Evolving Agent Skills: Feedback Dynamics over Multiple Rounds SetCon: Towards Open-Ended Referring Segmentation via Set-Level Concept Prediction
Reference 41
Source-reported events for the cited work
Unavailable: canonical work link unavailable.