Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-15T16:22:12.020753Z
Paper Citation Record · LEDGER
As of 15 August 2026, this Paper Citation Record lists 95 of 95 outbound references and 2 inbound Pith citation observations for arXiv:2509.06321.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-15T16:22:12.020753Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-15T06:32:42.880941+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-05-10T18:00:20.216268Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-05-11T10:11:04.772242Z
95 of 95 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 716f03b8-86ab-4e08-8c8d-74761c3720c9 · outbound
Text4Seg++: Advancing Image Segmentation via Generative Language Modeling A survey on multimodal large language models,
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0f62a146-a2f4-4d63-9bc6-f5571e8f2534 · outbound
Text4Seg++: Advancing Image Segmentation via Generative Language Modeling Visual instruction tuning,
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8dc1dd1d-1d75-4478-99a3-d81a985a0b0c · outbound
Text4Seg++: Advancing Image Segmentation via Generative Language Modeling Deepseek- vl: Towards real-world vision-language understanding,
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b1d28aa1-1afb-47d4-89ab-1b0f6dd3b58c · outbound
Text4Seg++: Advancing Image Segmentation via Generative Language Modeling Improved baselines with visual instruction tuning,
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4d2d77bb-cbc9-4ddb-b413-26d7105d731f · outbound
Text4Seg++: Advancing Image Segmentation via Generative Language Modeling Qwen-VL: A Versatile Vision-Language Model for Understanding, Localization, Text Reading, and Beyond
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5a9bf968-02f8-4a54-bb14-9d970c2c971a · outbound
Text4Seg++: Advancing Image Segmentation via Generative Language Modeling How far are we to gpt-4v? closing the gap to commercial multimodal models with open-source suites,
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation caee47b3-16b0-426b-83da-88d573f67b5c · outbound
Text4Seg++: Advancing Image Segmentation via Generative Language Modeling Moma: Multimodal llm adapter for fast personalized image generation,
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e6fbfd26-846e-440f-b710-2f666d48e2b6 · outbound
Text4Seg++: Advancing Image Segmentation via Generative Language Modeling Genartist: Multimodal llm as an agent for unified image generation and editing,
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f1d1bcbb-da5a-485d-80d1-2c248c5cca09 · outbound
Text4Seg++: Advancing Image Segmentation via Generative Language Modeling Visionllm: Large language model is also an open-ended decoder for vision-centric tasks,
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b331bf85-980c-48c2-9934-994d1a0e9ce7 · outbound
Text4Seg++: Advancing Image Segmentation via Generative Language Modeling Groma: Localized visual tokenization for grounding multimodal large language models,
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 94569e89-143d-482b-b03f-8cb59540ccbe · outbound
Text4Seg++: Advancing Image Segmentation via Generative Language Modeling Towards open vocabulary learning: A survey,
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8b45ef7e-c52e-4f57-bdae-c0cd4a2de9df · outbound
Text4Seg++: Advancing Image Segmentation via Generative Language Modeling NExT-Chat: An LMM for Chat, Detection and Segmentation
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c90b2773-5ec2-4b61-b6bc-3826548b16ba · outbound
Text4Seg++: Advancing Image Segmentation via Generative Language Modeling Transformer-based visual segmentation: A survey,
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4cae6a04-e525-4a39-8aec-5334820936e0 · outbound
Text4Seg++: Advancing Image Segmentation via Generative Language Modeling Smooseg: smoothness prior for unsupervised semantic segmentation,
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4c81b0c7-3da7-41b1-a837-d777d9576340 · outbound
Text4Seg++: Advancing Image Segmentation via Generative Language Modeling Lisa: Reasoning segmentation via large language model,
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 39fbcf90-08c0-4bca-a33e-f78a08036d04 · outbound
Text4Seg++: Advancing Image Segmentation via Generative Language Modeling Gsva: Generalized segmentation via multimodal large language models,
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 53d79d32-df32-4514-b455-bf21be807aa1 · outbound
Text4Seg++: Advancing Image Segmentation via Generative Language Modeling Groundhog: Grounding large language models to holistic segmentation,
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d4bfe972-1a09-4597-adb3-ad423ed28433 · outbound
Text4Seg++: Advancing Image Segmentation via Generative Language Modeling Multi-modal instruction tuned llms with fine-grained visual perception,
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation c682b7c6-6b0b-431b-b2b6-d712e43d3a44 · outbound
Text4Seg++: Advancing Image Segmentation via Generative Language Modeling Pixellm: Pixel reasoning with large multimodal model,
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8d824f5d-66f3-4eb5-920f-0e1ba053b51e · outbound
Text4Seg++: Advancing Image Segmentation via Generative Language Modeling Glamm: Pixel grounding large multimodal model,
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4b0231c8-78cc-47bb-bd3a-4d8fb6362b3f · outbound
Text4Seg++: Advancing Image Segmentation via Generative Language Modeling Segment anything,
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 2757456d-95a0-4a9d-8364-b9ccd471cf09 · outbound
Text4Seg++: Advancing Image Segmentation via Generative Language Modeling Jack of all tasks master of many: Designing general-purpose coarse-to-fine vision-language model,
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 8ce7c468-7c2d-4911-b0df-892351472ea5 · outbound
Text4Seg++: Advancing Image Segmentation via Generative Language Modeling Florence-2: Advancing a unified representation for a variety of vision tasks,
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5f83132b-68e7-42eb-a17b-282989b0cc3d · outbound
Text4Seg++: Advancing Image Segmentation via Generative Language Modeling Visionllm v2: An end-to-end generalist multimodal large language model for hundreds of vision-language tasks,
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 2c00a354-3c52-442e-95d3-202f7297586f · outbound
Text4Seg++: Advancing Image Segmentation via Generative Language Modeling Text4seg: Reimagining image segmentation as text generation,
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 3340fec4-8112-4d05-b5d3-b798fd76bfaf · outbound
Text4Seg++: Advancing Image Segmentation via Generative Language Modeling An image is worth 16x16 words: Transformers for image recognition at scale,
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 89d2949c-35d2-43a8-b989-53a6f514d1df · outbound
Text4Seg++: Advancing Image Segmentation via Generative Language Modeling Qwen2-VL: Enhancing Vision-Language Model's Perception of the World at Any Resolution
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a42fe197-fe8a-4913-9466-e31f781d7447 · outbound
Text4Seg++: Advancing Image Segmentation via Generative Language Modeling DeepSeek-VL2: Mixture-of-Experts Vision-Language Models for Advanced Multimodal Understanding
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 82cff4a6-f430-488c-b07e-fa39a01ab7d9 · outbound
Text4Seg++: Advancing Image Segmentation via Generative Language Modeling InternVL3: Exploring Advanced Training and Test-Time Recipes for Open-Source Multimodal Models
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation edecca24-37f4-4c2e-9ca0-c4797a88165d · outbound
Text4Seg++: Advancing Image Segmentation via Generative Language Modeling Flamingo: a visual language model for few-shot learning,
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 21d16457-5f00-4542-a3f3-967ad991e07b · outbound
Text4Seg++: Advancing Image Segmentation via Generative Language Modeling OpenFlamingo: An Open-Source Framework for Training Large Autoregressive Vision-Language Models
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 446592e4-4679-473f-b55a-3ee3f890c9fb · outbound
Text4Seg++: Advancing Image Segmentation via Generative Language Modeling Otter: A multi-modal model with in-context instruction tun- ing,
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b276f24d-9047-4ef1-8600-15cbdc992481 · outbound
Text4Seg++: Advancing Image Segmentation via Generative Language Modeling Blip-2: Bootstrapping language- image pre-training with frozen image encoders and large language models,
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 128d0749-aefc-4acd-ab02-6d2d38d36b82 · outbound
Text4Seg++: Advancing Image Segmentation via Generative Language Modeling InstructBLIP: Towards general-purpose vision-language models with instruction tuning,
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation dc0a4841-5d66-46b6-b9dd-a889b2aad7f5 · outbound
Text4Seg++: Advancing Image Segmentation via Generative Language Modeling mplug-owl2: Revolutionizing multi-modal large language model with modality collaboration,
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 05423cf7-ef37-4460-866a-4b857ceccefc · outbound
Text4Seg++: Advancing Image Segmentation via Generative Language Modeling Llava-next: Improved reasoning, ocr, and world knowledge,
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e7ebdc94-4c9d-4f1e-aaac-b5d4b732a407 · outbound
Text4Seg++: Advancing Image Segmentation via Generative Language Modeling Llava-uhd: an lmm perceiving any aspect ratio and high-resolution images,
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 3bd5248c-a49a-44c3-b423-d01028ee0af6 · outbound
Text4Seg++: Advancing Image Segmentation via Generative Language Modeling LLaV A-onevision: Easy visual task transfer,
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0524b036-e703-4ede-8226-be2ca8ba1850 · outbound
Text4Seg++: Advancing Image Segmentation via Generative Language Modeling Mini-Gemini: Mining the Potential of Multi-modality Vision Language Models
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f96255a2-8309-447a-b909-56648d063aca · outbound
Text4Seg++: Advancing Image Segmentation via Generative Language Modeling Monkey: Image resolution and text label are important things for large multi-modal models,
Reference 40
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation b9539a27-86bc-45bb-b126-6a766f17464a · outbound
Text4Seg++: Advancing Image Segmentation via Generative Language Modeling Sphinx: A mixer of weights, visual embeddings and image scales for multi-modal large language models,
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation f5fd926b-d934-461f-92c3-477900688a3a · outbound
Text4Seg++: Advancing Image Segmentation via Generative Language Modeling InstructSeg: Unifying Instructed Visual Segmentation with Multi-modal Large Language Models
Reference 42
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 45383fa1-5c15-439e-8906-6ef399568545 · outbound
Text4Seg++: Advancing Image Segmentation via Generative Language Modeling Expanding Performance Boundaries of Open-Source Multimodal Models with Model, Data, and Test-Time Scaling
Reference 43
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c6acb05d-fd47-4252-bbd5-d27fcc0f1210 · outbound
Text4Seg++: Advancing Image Segmentation via Generative Language Modeling Qwen2.5-VL Technical Report
Reference 44
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation dc969136-352c-41b6-b41a-a505387dc42b · outbound
Text4Seg++: Advancing Image Segmentation via Generative Language Modeling Omg-llava: Bridging image-level, object-level, pixel-level reasoning and understanding,
Reference 45
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 506b73e9-c295-4ebe-a4b0-bb695d08020b · outbound
Text4Seg++: Advancing Image Segmentation via Generative Language Modeling MMR: A large-scale benchmark dataset for multi-target and multi-granularity reasoning segmentation,
Reference 46
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation a4d57b4b-8fd7-4272-8936-30d94812745c · outbound
Text4Seg++: Advancing Image Segmentation via Generative Language Modeling SegLLM: Multi-round reasoning segmentation with large language models,
Reference 47
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 9e40fef4-cb57-4c04-9145-434de89ae1fb · outbound
Text4Seg++: Advancing Image Segmentation via Generative Language Modeling SegEarth-R1: Geospatial Pixel Reasoning via Large Language Model
Reference 48
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2d391006-c002-41de-887e-13e3b3027ba3 · outbound
Text4Seg++: Advancing Image Segmentation via Generative Language Modeling Geopix: A multimodal large language model for pixel-level image understanding in remote sensing,
Reference 49
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation b1685c66-070a-4431-85ea-9f62e4d035c8 · outbound
Text4Seg++: Advancing Image Segmentation via Generative Language Modeling Ufo: A unified approach to fine-grained visual perception via open-ended language interface,
Reference 50
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a11488e5-ef7a-44eb-8211-4e66ff77f7f4 · outbound
Text4Seg++: Advancing Image Segmentation via Generative Language Modeling Pixel-SAIL: Single Transformer For Pixel-Grounded Understanding
Reference 51
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 84c93dfc-4775-4d2c-845e-0d00910c8ed7 · outbound
Text4Seg++: Advancing Image Segmentation via Generative Language Modeling HiMTok: Learning Hierarchical Mask Tokens for Image Segmentation with Large Multimodal Model
Reference 52
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c98a7b17-c7fc-43fc-a776-a22a0b3c0603 · outbound
Text4Seg++: Advancing Image Segmentation via Generative Language Modeling Alto: Adaptive-length tokenizer for autoregressive mask generation,
Reference 53
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d15d53f2-8816-4c93-b271-92667ee7490e · outbound
Text4Seg++: Advancing Image Segmentation via Generative Language Modeling Pix2seq: A language modeling framework for object detection,
Reference 54
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation fe671d6a-c133-46d4-b53b-5648672120c8 · outbound
Text4Seg++: Advancing Image Segmentation via Generative Language Modeling Grounding dino: Marrying dino with grounded pre-training for open-set object detection,
Reference 55
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 306794c4-6ea0-41ed-8e0e-06b6f31e3e57 · outbound
Text4Seg++: Advancing Image Segmentation via Generative Language Modeling Universal instance perception as object discovery and retrieval,
Reference 56
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 0d5e6583-3433-414d-bfc1-87373ee463fe · outbound
Text4Seg++: Advancing Image Segmentation via Generative Language Modeling Grounding multimodal large language models to the world,
Reference 57
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 320582a0-6bac-4264-a043-f1fcc1cac410 · outbound
Text4Seg++: Advancing Image Segmentation via Generative Language Modeling Shikra: Unleashing Multimodal LLM's Referential Dialogue Magic
Reference 58
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e05b92de-6617-498c-99a7-d8d3fe58a783 · outbound
Text4Seg++: Advancing Image Segmentation via Generative Language Modeling Vitron: A unified pixel-level vision llm for understanding, generating, segmenting, edit- ing,
Reference 59
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 4b781d89-4d0c-4727-89d3-7ea2c2337776 · outbound
Text4Seg++: Advancing Image Segmentation via Generative Language Modeling Learning transferable visual models from natural language supervision,
Reference 60
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 230337c0-0811-4a7f-8cf2-040084193265 · outbound
Text4Seg++: Advancing Image Segmentation via Generative Language Modeling Sigmoid loss for language image pre-training,
Reference 61
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0bfcbda7-bbfb-401c-9a30-7e642fb9b842 · outbound
Text4Seg++: Advancing Image Segmentation via Generative Language Modeling Qwen3 Technical Report
Reference 62
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 84748385-428c-4c6c-a591-2c664d0d6fd0 · outbound
Text4Seg++: Advancing Image Segmentation via Generative Language Modeling LLaMA: Open and Efficient Foundation Language Models
Reference 63
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 52b70454-8e07-4d83-b61c-ef088c3260e4 · outbound
Text4Seg++: Advancing Image Segmentation via Generative Language Modeling Referitgame: Referring to objects in photographs of natural scenes,
Reference 64
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ef7d8813-6477-4dd3-88cf-18cd7bda100c · outbound
Text4Seg++: Advancing Image Segmentation via Generative Language Modeling Run-length encodings (corresp.),
Reference 65
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation f785469b-2cee-4162-ad49-b1de33aa02b5 · outbound
Text4Seg++: Advancing Image Segmentation via Generative Language Modeling LoRA: Low-rank adaptation of large language models,
Reference 66
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f5456534-717c-4bc7-bc20-d58eaba8d97c · outbound
Text4Seg++: Advancing Image Segmentation via Generative Language Modeling SAMRefiner: Taming segment anything model for universal mask refinement,
Reference 67
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation f3d74f23-94dc-4fbe-b79d-835510d30056 · outbound
Text4Seg++: Advancing Image Segmentation via Generative Language Modeling Swift: A scalable lightweight infrastructure for fine-tuning,
Reference 68
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 64db9197-ac38-4595-b229-491b711ab833 · outbound
Text4Seg++: Advancing Image Segmentation via Generative Language Modeling Decoupled weight decay regularization,
Reference 69
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d8d3f090-35ac-47ac-baae-11165d5b497a · outbound
Text4Seg++: Advancing Image Segmentation via Generative Language Modeling Zero: Memory optimizations toward training trillion parameter models,
Reference 70
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8c2f4997-aefa-4d50-9056-382aceb3729f · outbound
Text4Seg++: Advancing Image Segmentation via Generative Language Modeling Phrasecut: Language- based image segmentation in the wild,
Reference 71
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation e0ae129a-3705-4b2c-9108-52d019562233 · outbound
Text4Seg++: Advancing Image Segmentation via Generative Language Modeling Gres: Generalized referring expression segmentation,
Reference 72
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3686281b-e532-4e50-ac3c-a56d9b657180 · outbound
Text4Seg++: Advancing Image Segmentation via Generative Language Modeling Generation and comprehension of unambiguous object descriptions,
Reference 73
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 60c9acc2-87f3-4fc6-8bff-7b8dca37a979 · outbound
Text4Seg++: Advancing Image Segmentation via Generative Language Modeling Polyformer: Referring image segmentation as sequential polygon generation,
Reference 74
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 4247aed3-20bc-4b26-8943-9cf3d65b7d7e · outbound
Text4Seg++: Advancing Image Segmentation via Generative Language Modeling Language-aware vision transformer for referring segmentation,
Reference 75
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 234ead33-0ceb-4629-a3c0-72245f0cad71 · outbound
Text4Seg++: Advancing Image Segmentation via Generative Language Modeling Sam4mllm: Enhance multi-modal large language model for referring expression segmentation,
Reference 76
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 2c238a44-1cd0-4a52-8fb2-93a1ec3b2511 · outbound
Text4Seg++: Advancing Image Segmentation via Generative Language Modeling Segagent: Exploring pixel understanding capabilities in mllms by imitating human annotator trajectories,
Reference 77
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 6a763a57-4dd6-4ce9-84ff-dea2c2b65cc7 · outbound
Text4Seg++: Advancing Image Segmentation via Generative Language Modeling Popen: Preference-based optimization and ensemble for lvlm-based reasoning segmentation,
Reference 78
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation bef2a781-708f-493b-870f-7e3811d59447 · outbound
Text4Seg++: Advancing Image Segmentation via Generative Language Modeling Microsoft coco: Common objects in context,
Reference 79
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d3a068ef-ac45-4ca4-8037-8a2be6265799 · outbound
Text4Seg++: Advancing Image Segmentation via Generative Language Modeling Pix2Cap-COCO: Advancing Visual Comprehension via Pixel-Level Captioning
Reference 80
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8b1d6538-e8bc-4143-8f47-7199505655a3 · outbound
Text4Seg++: Advancing Image Segmentation via Generative Language Modeling Rotated multi-scale interaction network for referring remote sensing image seg- mentation,
Reference 81
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 17537b21-d030-4ca4-aeea-3d170b5374cb · outbound
Text4Seg++: Advancing Image Segmentation via Generative Language Modeling Seg-Zero: Reasoning-Chain Guided Segmentation via Cognitive Reinforcement
Reference 82
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0d2acdad-31d8-414b-8f75-d6c98ce3ada1 · outbound
Text4Seg++: Advancing Image Segmentation via Generative Language Modeling Clearclip: Decomposing clip representations for dense vision-language inference,
Reference 83
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation df2477ff-adf3-4639-ae1c-b2ed00f958c9 · outbound
Text4Seg++: Advancing Image Segmentation via Generative Language Modeling Proxyclip: Proxy attention improves clip for open-vocabulary segmentation,
Reference 84
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 9b5a401e-d20c-4b66-90d3-87f9f5ec3a74 · outbound
Text4Seg++: Advancing Image Segmentation via Generative Language Modeling Open-vocabulary universal image seg- mentation with maskclip,
Reference 85
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 978982d6-d188-4c57-8389-88d97b2b0a29 · outbound
Text4Seg++: Advancing Image Segmentation via Generative Language Modeling Groupvit: Semantic segmentation emerges from text supervision,
Reference 86
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 66daabaf-eb5a-4188-8002-d8627b4eda68 · outbound
Text4Seg++: Advancing Image Segmentation via Generative Language Modeling Open-vocabulary semantic segmentation with mask-adapted clip,
Reference 87
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation ed057f24-fdde-4764-90b4-5af354059273 · outbound
Text4Seg++: Advancing Image Segmentation via Generative Language Modeling San: side adapter network for open-vocabulary semantic segmentation,
Reference 88
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 1865d8d8-7f24-4bf2-bcee-fc1d0f0a6a95 · outbound
Text4Seg++: Advancing Image Segmentation via Generative Language Modeling LaSagnA: Language-based Segmentation Assistant for Complex Queries
Reference 89
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 087076ab-fc36-4a0d-8bcb-19179f2882ed · outbound
Text4Seg++: Advancing Image Segmentation via Generative Language Modeling GeoGround: A Unified Large Vision-Language Model for Remote Sensing Visual Grounding
Reference 90
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7cc02e5b-4d07-447e-b3c4-5230c55aa13b · outbound
Text4Seg++: Advancing Image Segmentation via Generative Language Modeling Semantic understanding of scenes through the ade20k dataset,
Reference 91
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a824abb5-81a9-416a-8517-7eccd65fa637 · outbound
Text4Seg++: Advancing Image Segmentation via Generative Language Modeling The role of context for object detection and semantic segmentation in the wild,
Reference 92
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9d36d4c7-c8c8-4889-bc75-68c29c1516b4 · outbound
Text4Seg++: Advancing Image Segmentation via Generative Language Modeling The pascal visual object classes challenge 2007,
Reference 93
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 01d8ea53-ab5f-4988-bb0a-26d37fc44b82 · outbound
Text4Seg++: Advancing Image Segmentation via Generative Language Modeling Side adapter network for open-vocabulary semantic segmentation,
Reference 94
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7ece75ed-0efc-4d9b-82aa-09cc53b2e085 · outbound
Text4Seg++: Advancing Image Segmentation via Generative Language Modeling Available: https://openreview.net/forum?id=lLmqxkfSIw
Reference 2024
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 76f9a169-93e0-42b8-a8fe-e1941f75b22b · inbound
RemoteAgent: Bridging Vague Human Intents and Earth Observation with RL-based Agentic MLLMs Text4Seg++: Advancing Image Segmentation via Generative Language Modeling
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 68a75981-2033-4edd-8651-8fe608d4a57d · inbound
LMMs Meet Object-Centric Vision: Understanding, Segmentation, Editing and Generation Text4Seg++: Advancing Image Segmentation via Generative Language Modeling
Reference 79
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.