Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-15T15:03:54.710735Z
Paper Citation Record · LEDGER
As of 16 August 2026, this Paper Citation Record lists 76 of 76 outbound references and 0 inbound Pith citation observations for arXiv:2608.02791.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-15T15:03:54.710735Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-15T06:32:42.880941+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
76 of 76 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 04b08a1d-1618-4655-a739-2cf0e76629af · outbound
Better, Stronger, Faster, and Broader: Structured All-Mask Prediction for MLLM-Based Segmentation Better, stronger, faster: Tackling the trilemma in mllm-based segmentation with simultaneous textual mask prediction,
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation b8bfe145-ca25-4b8d-8ae0-6e0548181006 · outbound
Better, Stronger, Faster, and Broader: Structured All-Mask Prediction for MLLM-Based Segmentation Qwen3 Technical Report
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3d1abe1e-a1fb-45e0-b662-4ffe1e1796eb · outbound
Better, Stronger, Faster, and Broader: Structured All-Mask Prediction for MLLM-Based Segmentation LION: Empowering multimodal large language model with dual-level visual knowledge,
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation ca71740f-6953-4b02-9706-29476b94628c · outbound
Better, Stronger, Faster, and Broader: Structured All-Mask Prediction for MLLM-Based Segmentation MLLMs know where to look: Training-free perception of small visual details with multimodal llms,
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation fd34f1a2-1f9d-4b9a-92b8-e75336cd8852 · outbound
Better, Stronger, Faster, and Broader: Structured All-Mask Prediction for MLLM-Based Segmentation Detect anything via next point prediction,
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ee3bf2eb-43a2-496a-881a-07ff99bee524 · outbound
Better, Stronger, Faster, and Broader: Structured All-Mask Prediction for MLLM-Based Segmentation InternVL3: Exploring Advanced Training and Test-Time Recipes for Open-Source Multimodal Models
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 43e2b31a-79ee-4d7c-a186-bdf99bf31cac · outbound
Better, Stronger, Faster, and Broader: Structured All-Mask Prediction for MLLM-Based Segmentation LLaVA-OneVision: Easy Visual Task Transfer
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c482a154-3187-4ee1-abb6-cc7a6053be76 · outbound
Better, Stronger, Faster, and Broader: Structured All-Mask Prediction for MLLM-Based Segmentation PhD: A ChatGPT-prompted visual halluci- nation evaluation dataset,
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation fe07297c-e113-4a74-a93a-53d6c32697a0 · outbound
Better, Stronger, Faster, and Broader: Structured All-Mask Prediction for MLLM-Based Segmentation Qwen2.5-VL Technical Report
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 74ed9fa0-c20d-4ae6-b070-15bd8cf4e1c8 · outbound
Better, Stronger, Faster, and Broader: Structured All-Mask Prediction for MLLM-Based Segmentation Segmentation as a plug-and-play capability for frozen multimodal LLMs,
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation a224531b-01d7-47f8-abe9-b5079e6bed0b · outbound
Better, Stronger, Faster, and Broader: Structured All-Mask Prediction for MLLM-Based Segmentation Text4Seg: Reimagining image segmentation as text generation,
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation cfd23eaa-e655-4fde-b097-85598dce54c2 · outbound
Better, Stronger, Faster, and Broader: Structured All-Mask Prediction for MLLM-Based Segmentation LISA: Reasoning segmentation via large language model,
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation e441a837-c859-4918-bd43-56a6b0bd4abb · outbound
Better, Stronger, Faster, and Broader: Structured All-Mask Prediction for MLLM-Based Segmentation GSV A: Generalized segmentation via multimodal large language models,
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation eb594ba4-8168-43c7-8986-3c49e5eeae95 · outbound
Better, Stronger, Faster, and Broader: Structured All-Mask Prediction for MLLM-Based Segmentation PixelLM: Pixel reasoning with large multimodal model,
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 7c0d53ec-299e-45d1-b124-5e19aee050d1 · outbound
Better, Stronger, Faster, and Broader: Structured All-Mask Prediction for MLLM-Based Segmentation MMR: A large-scale benchmark dataset for multi-target and multi- granularity reasoning segmentation,
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation fef7e871-c616-421b-9dec-eb69d416b313 · outbound
Better, Stronger, Faster, and Broader: Structured All-Mask Prediction for MLLM-Based Segmentation Reasoning to attend: Try to understand how<SEG>token works,
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 03e114b1-1560-4466-9e71-3d6dd8b1a0d0 · outbound
Better, Stronger, Faster, and Broader: Structured All-Mask Prediction for MLLM-Based Segmentation VisionLLM v2: An end-to-end generalist multimodal large language SUBMISSION 17 model for hundreds of vision-language tasks,
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation a2b4c67b-43b8-4186-bbfb-521c4f802bb2 · outbound
Better, Stronger, Faster, and Broader: Structured All-Mask Prediction for MLLM-Based Segmentation Seg-Zero: Reasoning-Chain Guided Segmentation via Cognitive Reinforcement
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1cd02046-2b11-4129-b374-725484920792 · outbound
Better, Stronger, Faster, and Broader: Structured All-Mask Prediction for MLLM-Based Segmentation SegAgent: Exploring pixel understanding capabilities in mllms by imitating human annotator trajectories,
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 406235f7-a893-424f-82dc-c6601765be84 · outbound
Better, Stronger, Faster, and Broader: Structured All-Mask Prediction for MLLM-Based Segmentation Text4Seg++: Advancing image segmentation via generative language modeling,
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 0593f58f-a32d-4410-bbc3-0c92be3785c8 · outbound
Better, Stronger, Faster, and Broader: Structured All-Mask Prediction for MLLM-Based Segmentation GLaMM: Pixel grounding large multimodal model,
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 2f93bd10-bd8f-48bc-8cda-0272fd706c1d · outbound
Better, Stronger, Faster, and Broader: Structured All-Mask Prediction for MLLM-Based Segmentation See say and segment: Teaching LMMs to overcome false premises,
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 6c2cd546-1986-45f4-9431-bc70bd7b7acd · outbound
Better, Stronger, Faster, and Broader: Structured All-Mask Prediction for MLLM-Based Segmentation VisionLLM: Large language model is also an open-ended decoder for vision-centric tasks,
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 29f44c12-a969-43fd-9612-dd62a074efcb · outbound
Better, Stronger, Faster, and Broader: Structured All-Mask Prediction for MLLM-Based Segmentation ReferItGame: Referring to objects in photographs of natural scenes,
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 4b02c12e-dc90-4336-ac4f-24207e27f9fb · outbound
Better, Stronger, Faster, and Broader: Structured All-Mask Prediction for MLLM-Based Segmentation Generation and comprehension of un- ambiguous object descriptions,
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 719a921b-2021-4c44-b8e0-e1ba6d91a1df · outbound
Better, Stronger, Faster, and Broader: Structured All-Mask Prediction for MLLM-Based Segmentation Qwen2-VL: Enhancing Vision-Language Model's Perception of the World at Any Resolution
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d939e361-afbb-4349-97ee-e866c4f81170 · outbound
Better, Stronger, Faster, and Broader: Structured All-Mask Prediction for MLLM-Based Segmentation Scaling Laws for Neural Language Models
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1426642a-64c0-4a48-85f4-5549ba467a59 · outbound
Better, Stronger, Faster, and Broader: Structured All-Mask Prediction for MLLM-Based Segmentation Hello GPT-4o,
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation e219408a-a3d9-46a8-9b7e-8476dfae5a2c · outbound
Better, Stronger, Faster, and Broader: Structured All-Mask Prediction for MLLM-Based Segmentation Google Gemini 2.5 Pro,
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 912d1f00-b430-4f06-84fd-bc4c4677137f · outbound
Better, Stronger, Faster, and Broader: Structured All-Mask Prediction for MLLM-Based Segmentation Empowering small VLMs to think with dynamic memorization and exploration,
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation ba301dd4-4fb2-48c0-9f4b-9cd6aa462077 · outbound
Better, Stronger, Faster, and Broader: Structured All-Mask Prediction for MLLM-Based Segmentation Flamingo: a visual language model for few-shot learning,
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4ef2b056-5ca2-4e0f-b851-d7cfdeda56ed · outbound
Better, Stronger, Faster, and Broader: Structured All-Mask Prediction for MLLM-Based Segmentation InstructBLIP: Towards general- purpose vision-language models with instruction tuning,
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation b1bfb3f6-1214-4997-bad2-880feadda6ea · outbound
Better, Stronger, Faster, and Broader: Structured All-Mask Prediction for MLLM-Based Segmentation Visual instruc- tion tuning,
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 4bcfd626-2366-4f5a-9e93-9a24e3ad82a9 · outbound
Better, Stronger, Faster, and Broader: Structured All-Mask Prediction for MLLM-Based Segmentation Improved baselines with visual instruction tuning,
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 601a7695-856f-42b3-9a2d-7cf4149bce0b · outbound
Better, Stronger, Faster, and Broader: Structured All-Mask Prediction for MLLM-Based Segmentation Qwen-VL: A Versatile Vision-Language Model for Understanding, Localization, Text Reading, and Beyond
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b60787d1-80fa-49d2-a7ce-2e81bbe1d8cf · outbound
Better, Stronger, Faster, and Broader: Structured All-Mask Prediction for MLLM-Based Segmentation Qwen3-VL Technical Report
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d7be95b9-76e9-4630-add3-b89b088f9cb8 · outbound
Better, Stronger, Faster, and Broader: Structured All-Mask Prediction for MLLM-Based Segmentation Latent Visual Reasoning
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cc610fa0-ba34-4b73-82ad-5953674b34df · outbound
Better, Stronger, Faster, and Broader: Structured All-Mask Prediction for MLLM-Based Segmentation Segment anything,
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 89d0900f-e339-48d5-9708-b9d2454358b2 · outbound
Better, Stronger, Faster, and Broader: Structured All-Mask Prediction for MLLM-Based Segmentation SegLLM: Multi- round reasoning segmentation with large language mod- els,
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 880f5ac2-2e34-495f-ab32-5e666196e4dd · outbound
Better, Stronger, Faster, and Broader: Structured All-Mask Prediction for MLLM-Based Segmentation MLLM can see? dynamic correction decoding for hallucination mitigation,
Reference 40
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 0dbd1d34-3a1f-4c57-a305-322c6a077190 · outbound
Better, Stronger, Faster, and Broader: Structured All-Mask Prediction for MLLM-Based Segmentation Grounding multimodal large lan- guage models to the world,
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 7a34b36a-3858-4006-9236-d7ef689eefbc · outbound
Better, Stronger, Faster, and Broader: Structured All-Mask Prediction for MLLM-Based Segmentation Shikra: Unleashing Multimodal LLM's Referential Dialogue Magic
Reference 42
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation aa5a3832-948a-4297-8a62-0e08b2f85793 · outbound
Better, Stronger, Faster, and Broader: Structured All-Mask Prediction for MLLM-Based Segmentation Perception tokens en- hance visual reasoning in multimodal language models,
Reference 43
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation eb3ae115-811f-4f68-b83c-eef6c98af6df · outbound
Better, Stronger, Faster, and Broader: Structured All-Mask Prediction for MLLM-Based Segmentation ViperGPT: Visual inference via python execution for reasoning,
Reference 44
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 0a5f2e30-0404-48a2-a1c9-a8cbc821d645 · outbound
Better, Stronger, Faster, and Broader: Structured All-Mask Prediction for MLLM-Based Segmentation Chameleon: Plug- and-play compositional reasoning with large language models,
Reference 45
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation bb0adc8a-f9b2-425d-908b-274ec596363f · outbound
Better, Stronger, Faster, and Broader: Structured All-Mask Prediction for MLLM-Based Segmentation MM-REACT: Prompting ChatGPT for Multimodal Reasoning and Action
Reference 46
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b7e4c956-5862-495f-832d-88bbaf4587e6 · outbound
Better, Stronger, Faster, and Broader: Structured All-Mask Prediction for MLLM-Based Segmentation Visual sketchpad: Sketching as a visual chain of thought for multimodal language models,
Reference 47
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 039fe8d4-6bfa-4653-aa56-f6504ca3128a · outbound
Better, Stronger, Faster, and Broader: Structured All-Mask Prediction for MLLM-Based Segmentation Machine Mental Imagery: Empower Multimodal Reasoning with Latent Visual Tokens
Reference 48
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d51d1135-2c12-457c-8f96-f903be002fff · outbound
Better, Stronger, Faster, and Broader: Structured All-Mask Prediction for MLLM-Based Segmentation Multimodal chain of continu- ous thought for latent-space reasoning in vision-language models,
Reference 49
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 11a287d2-81e2-49de-92d2-4c52baffc397 · outbound
Better, Stronger, Faster, and Broader: Structured All-Mask Prediction for MLLM-Based Segmentation GRES: Generalized referring expression segmentation,
Reference 50
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 47842052-bb3f-4e0b-83c1-8a29e5673b6b · outbound
Better, Stronger, Faster, and Broader: Structured All-Mask Prediction for MLLM-Based Segmentation Rotated multi-scale interaction network for referring remote sensing image segmentation,
Reference 51
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 1620cbd2-528e-4e73-91ae-c95901437297 · outbound
Better, Stronger, Faster, and Broader: Structured All-Mask Prediction for MLLM-Based Segmentation SegEarth-R1: Geospatial Pixel Reasoning via Large Language Model
Reference 52
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0b84a510-00d2-4954-b6bd-49f51a874a90 · outbound
Better, Stronger, Faster, and Broader: Structured All-Mask Prediction for MLLM-Based Segmentation COCO-Stuff: Thing and stuff classes in context,
Reference 53
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation d2717e80-0ea9-4c2e-9947-75f9d4fdbce6 · outbound
Better, Stronger, Faster, and Broader: Structured All-Mask Prediction for MLLM-Based Segmentation Microsoft COCO: Common objects in context,
Reference 54
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation eb93b69d-18f0-4682-9bec-984b6b5b5123 · outbound
Better, Stronger, Faster, and Broader: Structured All-Mask Prediction for MLLM-Based Segmentation Universal instance perception as object discovery and retrieval,
Reference 55
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 3ec1a4bd-74fd-4935-8aa8-974b566a410c · outbound
Better, Stronger, Faster, and Broader: Structured All-Mask Prediction for MLLM-Based Segmentation PolyFormer: Referring im- age segmentation as sequential polygon generation,
Reference 56
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 725c51a9-44cd-4e61-9b12-4e53da9d4a99 · outbound
Better, Stronger, Faster, and Broader: Structured All-Mask Prediction for MLLM-Based Segmentation Language-aware vision transformer for referring segmentation,
Reference 57
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 26b018bf-29de-478c-8b7d-71d677686b35 · outbound
Better, Stronger, Faster, and Broader: Structured All-Mask Prediction for MLLM-Based Segmentation Open-vocabulary semantic segmentation with mask-adapted clip,
Reference 58
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 803c6677-c8c7-45ba-a00e-0f9ba6204390 · outbound
Better, Stronger, Faster, and Broader: Structured All-Mask Prediction for MLLM-Based Segmentation NExT- Chat: An LMM for chat, detection and segmentation,
Reference 59
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 46fea211-ec73-42f4-af8c-c4d41d115b7f · outbound
Better, Stronger, Faster, and Broader: Structured All-Mask Prediction for MLLM-Based Segmentation GeoGround: A Unified Large Vision-Language Model for Remote Sensing Visual Grounding
Reference 60
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7e48c61d-4e4a-4342-99f8-ab0fd99b6040 · outbound
Better, Stronger, Faster, and Broader: Structured All-Mask Prediction for MLLM-Based Segmentation Semantic understanding of scenes through the ADE20K dataset,
Reference 61
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 181e0006-ebcc-47e4-bf79-0cb248aab9ad · outbound
Better, Stronger, Faster, and Broader: Structured All-Mask Prediction for MLLM-Based Segmentation The role of context for object detection and semantic segmentation in the wild,
Reference 62
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation eddf5a35-7074-4569-a919-967d4eb3d5a1 · outbound
Better, Stronger, Faster, and Broader: Structured All-Mask Prediction for MLLM-Based Segmentation The PASCAL visual object classes (VOC) challenge,
Reference 63
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation e1620c24-8c25-44af-a28b-225d4a69d47e · outbound
Better, Stronger, Faster, and Broader: Structured All-Mask Prediction for MLLM-Based Segmentation ClearCLIP: Decomposing clip representa- tions for dense vision-language inference,
Reference 64
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation dc9135cc-a1c3-4133-8670-8aa8dab1b5c1 · outbound
Better, Stronger, Faster, and Broader: Structured All-Mask Prediction for MLLM-Based Segmentation ProxyCLIP: Proxy attention improves clip for open-vocabulary segmentation,
Reference 65
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation ae84c898-2067-4b56-abdb-828bdc0b91e7 · outbound
Better, Stronger, Faster, and Broader: Structured All-Mask Prediction for MLLM-Based Segmentation Open-vocabulary universal image segmentation with maskclip,
Reference 66
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 284bd96e-62dd-4aec-8b06-0bf807a575c3 · outbound
Better, Stronger, Faster, and Broader: Structured All-Mask Prediction for MLLM-Based Segmentation GroupViT: Semantic segmenta- tion emerges from text supervision,
Reference 67
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation d0702a6f-98d4-4145-aafa-76ed3cdfd46c · outbound
Better, Stronger, Faster, and Broader: Structured All-Mask Prediction for MLLM-Based Segmentation SAN: Side adapter network for open-vocabulary semantic seg- mentation,
Reference 68
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 6bfc685c-e110-4a95-aa14-55fcf932c696 · outbound
Better, Stronger, Faster, and Broader: Structured All-Mask Prediction for MLLM-Based Segmentation LaSagnA: Language-based Segmentation Assistant for Complex Queries
Reference 69
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b02ad1b6-b270-415a-afc2-63967b3a525d · outbound
Better, Stronger, Faster, and Broader: Structured All-Mask Prediction for MLLM-Based Segmentation POPEN: Preference-based optimization and ensemble for LVLM-based reasoning segmentation,
Reference 70
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation f4f4dae6-1f65-4854-9c44-ec1f3a0942bc · outbound
Better, Stronger, Faster, and Broader: Structured All-Mask Prediction for MLLM-Based Segmentation MMMU: A massive multi-discipline multimodal understanding and reasoning benchmark for expert agi,
Reference 71
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 70d01109-c09b-4ca4-aea6-09f72332c59d · outbound
Better, Stronger, Faster, and Broader: Structured All-Mask Prediction for MLLM-Based Segmentation MMBench: Is your multi-modal model an all-around player?
Reference 72
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 93f5a652-b3c4-4b62-a4d2-0f96a4779fce · outbound
Better, Stronger, Faster, and Broader: Structured All-Mask Prediction for MLLM-Based Segmentation Are we on the right way for evaluating large vision-language models?
Reference 73
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9def69ac-b21b-49dc-9a2f-35e2016ac5a6 · outbound
Better, Stronger, Faster, and Broader: Structured All-Mask Prediction for MLLM-Based Segmentation Learn to explain: Multimodal reasoning via thought chains for science question answering,
Reference 74
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 2815211c-9384-4853-9633-3416355e1333 · outbound
Better, Stronger, Faster, and Broader: Structured All-Mask Prediction for MLLM-Based Segmentation Towards VQA models that can read,
Reference 75
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation f8b30706-2dfe-42d5-9f36-854c70f92195 · outbound
Better, Stronger, Faster, and Broader: Structured All-Mask Prediction for MLLM-Based Segmentation VizWiz grand challenge: Answering visual questions from blind people,
Reference 76
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
No inbound Pith citation observations are available.