Pith. sign in

Paper Citation Record · LEDGER

SynRES: Towards Referring Expression Segmentation in the Wild via Synthetic Data

As of 8 August 2026, this Paper Citation Record lists 60 of 60 outbound references and 0 inbound Pith citation observations for arXiv:2505.17695.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.17695 v1

Coverage vector

measured 60 of 60 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T14:47:27.120881Z

measured 60 of 60 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

60 of 60 outbound references displayed

  • verified exact3
  • verified fuzzy49
  • unresolved8
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 18aa0765-f4fc-43f3-9784-94f3d02bdcaf · outbound

This paper cites GPT-4 Technical Report.

SynRES: Towards Referring Expression Segmentation in the Wild via Synthetic Data GPT-4 Technical Report

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-07T14:47:20.053296Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:47:20.053296Z digest=sha256:18449f621ca039ba5c5bb673bcda17a24a3cc08e50f21f44c57fb12e8ceeeb19

Observation db9ff230-e126-4178-a790-947e329204ee · outbound

This paper cites Vqa: Visual question answering.

SynRES: Towards Referring Expression Segmentation in the Wild via Synthetic Data Vqa: Visual question answering

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:47:40.174662Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:47:20.224226Z digest=sha256:f79c2ad67f6d7fe6faba5c2908f57da39a21bf69b0df7954ddd2c3f400ef2c20

Observation 75d2d425-9aed-44f3-bff1-77c513afcb5e · outbound

This paper cites Coco- stuff: Thing and stuff classes in context.

SynRES: Towards Referring Expression Segmentation in the Wild via Synthetic Data Coco- stuff: Thing and stuff classes in context

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:47:39.965237Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:47:20.354744Z digest=sha256:ee8cd4a9302330dc5a466cce52ac43a393cf0f7bdaee598a3d37e9273d384c15

Observation 0040dbe8-b07a-4196-b433-0c795bcf9153 · outbound

This paper cites Detect what you can: De- tecting and representing objects using holistic models and body parts.

SynRES: Towards Referring Expression Segmentation in the Wild via Synthetic Data Detect what you can: De- tecting and representing objects using holistic models and body parts

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:47:39.855083Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:47:20.582089Z digest=sha256:3919e403ddb44ab30578e00708de7baa01fcd6012de44121027359c5dbdb8f62

Observation 29e63728-7e7c-4d1d-840a-d251c80e6c04 · outbound

This paper cites Sam4mllm: Enhance multi- modal large language model for referring expression seg- mentation.

SynRES: Towards Referring Expression Segmentation in the Wild via Synthetic Data Sam4mllm: Enhance multi- modal large language model for referring expression seg- mentation

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:47:39.662221Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:47:20.719249Z digest=sha256:61522756d3d5086223b4923dabe04dccbcc6a31b872c31064e8adb68695fcd2e

Observation 593d3b04-b604-4091-b064-ecf33998db1b · outbound

This paper cites The cityscapes dataset for semantic urban scene understanding.

SynRES: Towards Referring Expression Segmentation in the Wild via Synthetic Data The cityscapes dataset for semantic urban scene understanding

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:47:39.388987Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:47:20.814327Z digest=sha256:46be9d89e59ab97b3c6f19fd5906e2be6e93f36927435fbe3d18a2348ebc8de1

Observation ad7a91b9-c475-43c5-a60d-ee70e36b05ab · outbound

This paper cites Divergen: Improving instance segmentation by learning wider data distribution with more diverse generative data.

SynRES: Towards Referring Expression Segmentation in the Wild via Synthetic Data Divergen: Improving instance segmentation by learning wider data distribution with more diverse generative data

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:47:39.171952Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:47:20.945150Z digest=sha256:a14488289c0967a1f53101bcfc2d332ce3dcff6a460f4c96011f366ee1cfc358

Observation 4f51af55-ece7-4008-92f8-d16910be9951 · outbound

This paper cites Finding nemo: Negative- mined mosaic augmentation for referring image segmentation.

SynRES: Towards Referring Expression Segmentation in the Wild via Synthetic Data Finding nemo: Negative- mined mosaic augmentation for referring image segmentation

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:47:38.874547Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:47:21.055373Z digest=sha256:fa888434e62d1550fbf61fd87a76a8a18b5bdde6e9765ceb6e706afe21ffd07c

Observation ed78cdbe-d00e-4fc0-ba65-b58e6c349938 · outbound

This paper cites Mixgen: A new multi- modal data augmentation.

SynRES: Towards Referring Expression Segmentation in the Wild via Synthetic Data Mixgen: A new multi- modal data augmentation

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:47:38.621618Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:47:21.156345Z digest=sha256:8cb211360c1084a99f52cb0078e2b10d8a3681e777086c93ebc32058b18b283c

Observation 6b5d6fab-f880-40aa-ab36-6cfdcd249263 · outbound

This paper cites Partimagenet: A large, high-quality dataset of parts.

SynRES: Towards Referring Expression Segmentation in the Wild via Synthetic Data Partimagenet: A large, high-quality dataset of parts

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:47:38.344668Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:47:21.270993Z digest=sha256:14096b99a0300742d7aee4c71bc0170d938955ceb45d94cbb296f9394b3a5163

Observation 4d663812-4d0e-492c-a258-9c9a1b0d5a54 · outbound

This paper cites Denoising dif- fusion probabilistic models.

SynRES: Towards Referring Expression Segmentation in the Wild via Synthetic Data Denoising dif- fusion probabilistic models

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-07T14:47:21.417306Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:47:21.417306Z digest=sha256:50a5cd5befb60d1beb024f8b58c2ea939e633c5993803f15475d64bbcbbf11df

Observation cc5a2dd0-2c42-4778-800d-c7a349015700 · outbound

This paper cites Beyond one-to-one: Rethinking the referring image segmentation.

SynRES: Towards Referring Expression Segmentation in the Wild via Synthetic Data Beyond one-to-one: Rethinking the referring image segmentation

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:47:38.073131Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:47:21.520435Z digest=sha256:5fab751da7ec6f0cfc202ce9be1a0efb1eeea72727f54f4745bc3a487ddf4fd5

Observation 982feeb2-8e80-48da-ad7e-2911838a7c6b · outbound

This paper cites GPT-4o System Card.

SynRES: Towards Referring Expression Segmentation in the Wild via Synthetic Data GPT-4o System Card

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-07T14:47:21.632881Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:47:21.632881Z digest=sha256:f986a9f6cdd97095d7b65f46a1f1904981842a76e1c9815cf266505c164d1b6a

Observation 04aa2b42-bb28-45f9-991e-09beeba581d7 · outbound

This paper cites ARMADA: Attribute-Based Multimodal Data Augmentation.

SynRES: Towards Referring Expression Segmentation in the Wild via Synthetic Data ARMADA: Attribute-Based Multimodal Data Augmentation

Reference 14

Resolution
verified exact
local_arxiv, observed 2026-08-07T14:47:27.779282Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:47:21.743339Z digest=sha256:12c5b381847ba3b45d24411bd54309692e059f314abf0010459440e120c8bd67

Observation 81d70f02-e4d9-499e-a0d0-6180e3306afe · outbound

This paper cites Referitgame: Referring to objects in pho- tographs of natural scenes.

SynRES: Towards Referring Expression Segmentation in the Wild via Synthetic Data Referitgame: Referring to objects in pho- tographs of natural scenes

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:47:37.826592Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:47:21.838974Z digest=sha256:16ef735a2d7bf7ab04b1d1a93eddb06e75852804e49def17b3b097a963c46e83

Observation ae266e83-b34f-4ec4-97a8-e409dbdb1892 · outbound

This paper cites Segment any- thing.

SynRES: Towards Referring Expression Segmentation in the Wild via Synthetic Data Segment any- thing

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:47:37.601437Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:47:21.921379Z digest=sha256:8ad4712c55389ed701cc11c331753b4393228f06049bfefe8f6dd8b638095e14

Observation 1bacd787-1def-49b8-b1b3-46cb72466998 · outbound

This paper cites Lisa: Reasoning segmentation via large language model.

SynRES: Towards Referring Expression Segmentation in the Wild via Synthetic Data Lisa: Reasoning segmentation via large language model

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:47:37.363436Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:47:22.006435Z digest=sha256:a4498dd85d49630664050c8febf5beedc3160a57d21105b53c698c3ea7c21bc5

Observation 698c7ee5-b400-4e3a-9262-594cd968a8ac · outbound

This paper cites Bigdatasetgan: Synthe- sizing imagenet with pixel-wise annotations.

SynRES: Towards Referring Expression Segmentation in the Wild via Synthetic Data Bigdatasetgan: Synthe- sizing imagenet with pixel-wise annotations

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:47:37.057411Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:47:22.120552Z digest=sha256:81de38f6e0008d9d8685d87f2b6aeb46c5e1259d43e8426c379b485f7a3cded6

Observation b7c1f8cc-3b97-41dd-b890-b773d04e95ec · outbound

This paper cites Gres: Gener- alized referring expression segmentation.

SynRES: Towards Referring Expression Segmentation in the Wild via Synthetic Data Gres: Gener- alized referring expression segmentation

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:47:36.829544Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:47:22.239095Z digest=sha256:a73c3ad5dfbb438ad969ece7e48613782683b5003540251a6a107634ff0285f5

Observation f3e6ca16-c619-4cac-8b4d-c154571f153c · outbound

This paper cites Improved baselines with visual instruction tuning.

SynRES: Towards Referring Expression Segmentation in the Wild via Synthetic Data Improved baselines with visual instruction tuning

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:47:36.591923Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:47:22.322938Z digest=sha256:5263c9cc7d0133f96e80c6b6fd42987cf99c389a291978e84b8b3b2907220656

Observation e1829636-215f-4de0-8a45-31ebf64feeab · outbound

This paper cites Visual instruction tuning.

SynRES: Towards Referring Expression Segmentation in the Wild via Synthetic Data Visual instruction tuning

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:47:36.367977Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:47:22.405143Z digest=sha256:0a58268bb43a264dbd86c46475dd0c26110bebdc2d8393dc2ca83e146bb72dc3

Observation 7b0b0c59-ebbc-4ddc-bb91-8141aea95786 · outbound

This paper cites Learning multimodal data augmentation in feature space.

SynRES: Towards Referring Expression Segmentation in the Wild via Synthetic Data Learning multimodal data augmentation in feature space

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:47:36.085520Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:47:22.497171Z digest=sha256:4a0d2f94733d7f131e0e86f5e617364c52917554a1e3e8fe2a365390fa400821

Observation 2b756d41-74bb-40b7-8fa0-0565df969ac0 · outbound

This paper cites Generation and com- prehension of unambiguous object descriptions.

SynRES: Towards Referring Expression Segmentation in the Wild via Synthetic Data Generation and com- prehension of unambiguous object descriptions

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:47:35.769701Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:47:22.582842Z digest=sha256:1016763074406537c25a0513089cd907c392bc37375e6b45bc13d8527c577bc6

Observation 2d0486e6-a933-42fd-a5ad-faba589ad9f0 · outbound

This paper cites Arm- bench: An object-centric benchmark dataset for robotic ma- nipulation.

SynRES: Towards Referring Expression Segmentation in the Wild via Synthetic Data Arm- bench: An object-centric benchmark dataset for robotic ma- nipulation

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:47:35.465501Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:47:22.655738Z digest=sha256:e65eb0be4860bf8b30192501dafb01a17f981642ffb0ba3a2584aed44c33c3f5

Observation 376cf2ff-cf1e-4986-b3dc-81d7c8a7364a · outbound

This paper cites The mapillary vistas dataset for semantic understanding of street scenes.

SynRES: Towards Referring Expression Segmentation in the Wild via Synthetic Data The mapillary vistas dataset for semantic understanding of street scenes

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:47:35.253944Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:47:22.831522Z digest=sha256:042faa2b9ec91a0998d4059c77f14d992814a1377c47df468b3f38861256ac81

Observation 8c7a92a1-16bf-4206-a287-7634dca742da · outbound

This paper cites Dataset diffusion: Diffusion-based synthetic data generation for pixel-level semantic segmentation.

SynRES: Towards Referring Expression Segmentation in the Wild via Synthetic Data Dataset diffusion: Diffusion-based synthetic data generation for pixel-level semantic segmentation

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:47:34.982424Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:47:22.906508Z digest=sha256:9636bb8aae6a8ce991650a64789c25336b01a7e897086c971593f5904eaebd45

Observation 6e601524-4404-4457-948b-7ebddc73b350 · outbound

This paper cites Learning transferable visual models from natural language supervision.

SynRES: Towards Referring Expression Segmentation in the Wild via Synthetic Data Learning transferable visual models from natural language supervision

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:47:34.697135Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:47:23.020067Z digest=sha256:18b5c3062fe69923851da89e3b21d2908addf98a78bcf21ba74f3492441d1aac

Observation e2e5d3db-c511-4199-8a54-e25671517fc1 · outbound

This paper cites Paco: Parts and attributes of common objects.

SynRES: Towards Referring Expression Segmentation in the Wild via Synthetic Data Paco: Parts and attributes of common objects

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:47:34.330678Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:47:23.143071Z digest=sha256:f937a8f02bf6c8ff53609dd41874e668ee612a475bb338727811c4288b988bc7

Observation fa6d6adf-7a56-43be-8858-94a792959ffd · outbound

This paper cites Glamm: Pixel grounding large multimodal model.

SynRES: Towards Referring Expression Segmentation in the Wild via Synthetic Data Glamm: Pixel grounding large multimodal model

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:47:34.086363Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:47:23.247935Z digest=sha256:8f81fde485727d30f7d2457dea8b8bfa14554d2e3582522acc1c166c8ea76566

Observation 86e7719c-f194-4a07-ab95-adc4a619e2ed · outbound

This paper cites SAM 2: Segment anything in images and videos.

SynRES: Towards Referring Expression Segmentation in the Wild via Synthetic Data SAM 2: Segment anything in images and videos

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:47:33.780743Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:47:23.370944Z digest=sha256:b11f3b746fc35a55a21e31056bc5fbd9cdad978ae94c220e5f594c9d993d94af

Observation c4b4b0f9-a4a7-4e5d-bbd5-950cb452100a · outbound

This paper cites Pixellm: Pixel reasoning with large multimodal model.

SynRES: Towards Referring Expression Segmentation in the Wild via Synthetic Data Pixellm: Pixel reasoning with large multimodal model

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:47:33.485800Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:47:23.476000Z digest=sha256:6ad448f11fa2ad0b282d158cabcd2e320de898ba7215bfccbc311c9c3745c91b

Observation 9cb64232-9601-4aaa-8e40-c94bc1fd775a · outbound

This paper cites Grounding of textual phrases in images by reconstruction.

SynRES: Towards Referring Expression Segmentation in the Wild via Synthetic Data Grounding of textual phrases in images by reconstruction

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:47:33.238447Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:47:23.585827Z digest=sha256:6d333220e16031837ffe67a0f9ab4da434128eee887ab9c7031863042c45bcbc

Observation a7751b07-0325-4461-8730-f2f0486faf37 · outbound

This paper cites CrowdHuman: A Benchmark for Detecting Human in a Crowd.

SynRES: Towards Referring Expression Segmentation in the Wild via Synthetic Data CrowdHuman: A Benchmark for Detecting Human in a Crowd

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-07T14:47:23.699161Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:47:23.699161Z digest=sha256:47a131d5be0f5497b88417864c04a94d2b864b120fbf1de38eb8927400f886cb

Observation a09e70b1-3ba4-4121-a076-ae22b72a40aa · outbound

This paper cites Denoising diffusion implicit models.

SynRES: Towards Referring Expression Segmentation in the Wild via Synthetic Data Denoising diffusion implicit models

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:47:33.015018Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:47:23.808606Z digest=sha256:837113542680db8a645a9d2d290f9da449fb82c32247bb0ed36dcecba6dbbc82

Observation 5f946cb7-fbce-46cf-8317-c2e35efda692 · outbound

This paper cites DreamMask: Boosting Open-vocabulary Panoptic Segmentation with Synthetic Data.

SynRES: Towards Referring Expression Segmentation in the Wild via Synthetic Data DreamMask: Boosting Open-vocabulary Panoptic Segmentation with Synthetic Data

Reference 35

Resolution
verified exact
local_arxiv, observed 2026-08-07T14:47:27.583189Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:47:23.916439Z digest=sha256:5b5571a2b93bf03b2721c510fbbcd519acd977f8f6a8ca156acca51e3dee7f3c

Observation 110b8971-dade-48ef-83e0-968aa9e3a808 · outbound

This paper cites Cris: Clip-driven referring image segmentation.

SynRES: Towards Referring Expression Segmentation in the Wild via Synthetic Data Cris: Clip-driven referring image segmentation

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:47:32.788625Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:47:24.069836Z digest=sha256:0377e39c3210a3b76c9968a0468c2443fe17790d2a78430dd0e0c713d32ecac9

Observation af2dc787-bdba-49e3-aa9e-79ca09ac039e · outbound

This paper cites Towards reporting bias in visual-language datasets: bimodal augmentation by decoupling object-attribute association.

SynRES: Towards Referring Expression Segmentation in the Wild via Synthetic Data Towards reporting bias in visual-language datasets: bimodal augmentation by decoupling object-attribute association

Reference 37

Resolution
verified exact
local_arxiv, observed 2026-08-07T14:47:27.366597Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:47:24.223505Z digest=sha256:5b6cc7815ee11ad08af50a18a528ed9f36a594d9f346cc6286a0744cf55d949c

Observation 311c22a3-2dc1-427d-811d-a8c2a2111c5b · outbound

This paper cites Diffumask: Synthesizing images with pixel-level annotations for semantic segmentation using diffu- sion models.

SynRES: Towards Referring Expression Segmentation in the Wild via Synthetic Data Diffumask: Synthesizing images with pixel-level annotations for semantic segmentation using diffu- sion models

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:47:32.549960Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:47:24.377245Z digest=sha256:bae4a3f8b46627283ae5d4ca78d8995be88894e81f5cffc7d7e66fec259fa8cf

Observation 65cc8935-735d-46fc-a756-96c111073256 · outbound

This paper cites Gsva: Generalized segmentation via multimodal large language models.

SynRES: Towards Referring Expression Segmentation in the Wild via Synthetic Data Gsva: Generalized segmentation via multimodal large language models

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:47:32.290075Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:47:24.527351Z digest=sha256:90f872255d4a59cd39a675420d32ae183c134082566a0deafe542433b71f8bc2

Observation 664af2dd-7586-4e2f-b56f-a4755bf3bebb · outbound

This paper cites SANA: Efficient High-Resolution Image Synthesis with Linear Diffusion Transformers.

SynRES: Towards Referring Expression Segmentation in the Wild via Synthetic Data SANA: Efficient High-Resolution Image Synthesis with Linear Diffusion Transformers

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-07T14:47:24.674218Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:47:24.674218Z digest=sha256:62e170c6df617cfe0853b49e2fb42e621d93f67419eefaa26da16f86a078cec6

Observation f9c2fbe1-7532-40b5-a8b8-a64bd7d7408e · outbound

This paper cites Mosaicfusion: Diffusion models as data augmenters for large vocabulary instance segmentation.

SynRES: Towards Referring Expression Segmentation in the Wild via Synthetic Data Mosaicfusion: Diffusion models as data augmenters for large vocabulary instance segmentation

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:47:32.060147Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:47:24.804534Z digest=sha256:97f378d9299c07a947a01515aff6d655c89fbda22d916556e8955428549459fc

Observation 896b32b1-bbf4-4502-85ab-79dc11b0e4f7 · outbound

This paper cites Bridging vision and language encoders: Parameter-efficient tuning for referring image segmentation.

SynRES: Towards Referring Expression Segmentation in the Wild via Synthetic Data Bridging vision and language encoders: Parameter-efficient tuning for referring image segmentation

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:47:31.877017Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:47:24.935045Z digest=sha256:4ddc3307b04f3741a4c529379dc8eddf7b751a1f4d81ffce3cb403a957cea1b9

Observation d107d3a5-db7d-4cc3-b2c0-5f562362d384 · outbound

This paper cites Panoptic scene graph gen- eration.

SynRES: Towards Referring Expression Segmentation in the Wild via Synthetic Data Panoptic scene graph gen- eration

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:47:31.620806Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:47:25.078638Z digest=sha256:ac91a11d376f1d031d8d765d345097db5db8b31efcdef12c8552b089b526b999

Observation 1365327d-86c0-4d4a-b647-c82600da3a52 · outbound

This paper cites Freemask: Synthetic images with dense annotations make stronger segmentation models.

SynRES: Towards Referring Expression Segmentation in the Wild via Synthetic Data Freemask: Synthetic images with dense annotations make stronger segmentation models

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:47:31.349947Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:47:25.211282Z digest=sha256:dea7d3b5fb08a72124b5c7c3754077224b86b8ca470e53eb78d7076ee15da3df

Observation b8f52a39-dab2-41d0-8719-3e2d99ac9dbc · outbound

This paper cites Lavt: Language-aware vi- sion transformer for referring image segmentation.

SynRES: Towards Referring Expression Segmentation in the Wild via Synthetic Data Lavt: Language-aware vi- sion transformer for referring image segmentation

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:47:31.049158Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:47:25.303573Z digest=sha256:c856aea3ac5cd1c25c166319dc40d037891a795d77a16f15d1e42548837da342

Observation 5b90db18-c264-45a0-a000-dc216eddebb7 · outbound

This paper cites Seggen: Supercharging segmentation models with text2mask and mask2img synthesis.

SynRES: Towards Referring Expression Segmentation in the Wild via Synthetic Data Seggen: Supercharging segmentation models with text2mask and mask2img synthesis

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:47:30.759804Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:47:25.460160Z digest=sha256:9648626bd710e59c92b53de5c5f4325dbc8d1e039520a3b4d7f7357eb175d6c5

Observation 9e21ba7b-94d9-4cd1-8b4c-13e978fbf85c · outbound

This paper cites Pix2Cap-COCO: Advancing Visual Comprehension via Pixel-Level Captioning.

SynRES: Towards Referring Expression Segmentation in the Wild via Synthetic Data Pix2Cap-COCO: Advancing Visual Comprehension via Pixel-Level Captioning

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-07T14:47:25.624443Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:47:25.624443Z digest=sha256:cecbdbe5f3daebe8a21190fbf33231eeeb808d640f3865c7e143bdcee448cb64

Observation be04d225-2071-4d7c-9095-03a4d872127f · outbound

This paper cites From image descriptions to visual denotations: New similarity metrics for semantic inference over event descrip- tions.

SynRES: Towards Referring Expression Segmentation in the Wild via Synthetic Data From image descriptions to visual denotations: New similarity metrics for semantic inference over event descrip- tions

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:47:30.484684Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:47:25.738988Z digest=sha256:fbb74f189b70484ac94694a46af2106bcf7a5028fa7856edfa34faf5b347be27

Observation acee422e-2b11-4869-b89b-d5faaecd52b9 · outbound

This paper cites Coca: Contrastive captioners are image-text foundation models.

SynRES: Towards Referring Expression Segmentation in the Wild via Synthetic Data Coca: Contrastive captioners are image-text foundation models

Reference 49

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:47:30.246850Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:47:25.851673Z digest=sha256:b3872f24e1d9c011d28e21e2d050071be34f8179e49c083ae6b3e3d985ba512f

Observation 298725ce-b1b7-465e-9404-4ce1c629a6b5 · outbound

This paper cites Modeling context in referring expressions.

SynRES: Towards Referring Expression Segmentation in the Wild via Synthetic Data Modeling context in referring expressions

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:47:30.004079Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:47:25.960341Z digest=sha256:2d09c64c82bdcad1dfd9a9df893b482772fa0045a443b4c2f62207f786124c7e

Observation 53ee01ac-5595-4c72-ad2d-3706a951b1f9 · outbound

This paper cites Pseudo- ris: Distinctive pseudo-supervision generation for referring image segmentation.

SynRES: Towards Referring Expression Segmentation in the Wild via Synthetic Data Pseudo- ris: Distinctive pseudo-supervision generation for referring image segmentation

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:47:29.677320Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:47:26.080046Z digest=sha256:572cb37497ee39e27a01f615c05c0e48e7f4a7fd20b47f353f2a57e437a49762

Observation 029d00c9-a3b2-46ca-a571-56dae9d18660 · outbound

This paper cites Revisiting counterfactual prob- lems in referring expression comprehension.

SynRES: Towards Referring Expression Segmentation in the Wild via Synthetic Data Revisiting counterfactual prob- lems in referring expression comprehension

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:47:29.418166Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:47:26.169375Z digest=sha256:98b916a8e5a773442ff0d249caddd4d67583c509909a7ac2a83a8c208fc951c0

Observation f413083f-ea99-4897-bd65-4ed33e1fddfe · outbound

This paper cites Datasetgan: Efficient labeled data factory with minimal human effort.

SynRES: Towards Referring Expression Segmentation in the Wild via Synthetic Data Datasetgan: Efficient labeled data factory with minimal human effort

Reference 53

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:47:29.165630Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:47:26.285795Z digest=sha256:db6e59c341c9d6c3a765f41970cefaa3340207f69bdbe0d0a03e64ba03f2c38f

Observation eea956fa-8b55-47b8-904c-11aabd935bef · outbound

This paper cites EVF-SAM: Early Vision-Language Fusion for Text-Prompted Segment Anything Model.

SynRES: Towards Referring Expression Segmentation in the Wild via Synthetic Data EVF-SAM: Early Vision-Language Fusion for Text-Prompted Segment Anything Model

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-07T14:47:26.416689Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:47:26.416689Z digest=sha256:f9b1812444355247d39c962180b9b288e22a41beb6b3eef23d989a7ecec642eb

Observation ca4ff209-3265-47d3-8c01-6989a655964e · outbound

This paper cites Psalm: Pixelwise segmentation with large multi-modal model.

SynRES: Towards Referring Expression Segmentation in the Wild via Synthetic Data Psalm: Pixelwise segmentation with large multi-modal model

Reference 55

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:47:28.928964Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:47:26.551994Z digest=sha256:417a8b273488fcc9809f7a06e6f02ee45544b48bafac1607fc7cb8d97781671b

Observation f7082026-268a-4aff-b4bf-f8b162d5061d · outbound

This paper cites X-paste: Revisiting scalable copy-paste for in- stance segmentation using clip and stablediffusion.

SynRES: Towards Referring Expression Segmentation in the Wild via Synthetic Data X-paste: Revisiting scalable copy-paste for in- stance segmentation using clip and stablediffusion

Reference 56

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:47:28.685348Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:47:26.658210Z digest=sha256:9be84fd1400dcca7b3c9af2a233c8af265604e527c812d993b4f83fb784ac845

Observation faaab8c1-8e33-4b9d-bde0-3cda016368a8 · outbound

This paper cites Unleashing text-to-image diffusion models for visual perception.

SynRES: Towards Referring Expression Segmentation in the Wild via Synthetic Data Unleashing text-to-image diffusion models for visual perception

Reference 57

Resolution
unresolved
no resolver link, observed 2026-08-07T14:47:26.760787Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:47:26.760787Z digest=sha256:05ef064b0ffdcc6a8b9df4c603b6dfb54e812740f031df00bce16e7ca8b2cc2a

Observation 7a5485ab-0c7b-4bbc-a4f3-e759d859ec47 · outbound

This paper cites Scene parsing through ade20k dataset.

SynRES: Towards Referring Expression Segmentation in the Wild via Synthetic Data Scene parsing through ade20k dataset

Reference 58

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:47:28.461809Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:47:26.890631Z digest=sha256:76fca23cacd7d690d988c3ba6e32698a0c232620e9ac0d40ffb2e8ec6a3fe7f0

Observation a9a2d513-693c-4783-b193-ac6be2c89da3 · outbound

This paper cites Generalized decoding for pixel, image, and language.

SynRES: Towards Referring Expression Segmentation in the Wild via Synthetic Data Generalized decoding for pixel, image, and language

Reference 59

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:47:28.237992Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:47:27.016843Z digest=sha256:9e5ee33dfad03f688d471f4d5f50c92db718e66336f0fcc182dd6cd79d6f8717

Observation 926b136e-3d50-499b-99db-ad42737ca299 · outbound

This paper cites the cat sitting on the bench next to big green wooden boat in the center of the image.

SynRES: Towards Referring Expression Segmentation in the Wild via Synthetic Data the cat sitting on the bench next to big green wooden boat in the center of the image

Reference 60

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:47:27.953849Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:47:27.120881Z digest=sha256:1e360617c0f7e684a85e49e9f4e472425890da3fc4287950e3c3d32ed1182484

Pith citing papers

No inbound Pith citation observations are available.