Pith. sign in

Paper Citation Record · LEDGER

SynRES: Towards Referring Expression Segmentation in the Wild via Synthetic Data

As of 9 August 2026, this Paper Citation Record lists 60 of 60 outbound references and 0 inbound Pith citation observations for arXiv:2505.17695.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.17695 v1

Coverage vector

measured 60 of 60 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T14:47:27.120881Z

measured 60 of 60 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

60 of 60 outbound references displayed

  • verified exact3
  • verified fuzzy49
  • unresolved8
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 18aa0765-f4fc-43f3-9784-94f3d02bdcaf · outbound

This paper cites GPT-4 Technical Report.

SynRES: Towards Referring Expression Segmentation in the Wild via Synthetic Data GPT-4 Technical Report

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-07T14:47:20.053296Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:47:20.053296Z digest=sha256:7807aed0a9f438c6d0b20d7ae6fbd4f47961d8c7c09c6b0a5c3faae3924bc2cb

Observation db9ff230-e126-4178-a790-947e329204ee · outbound

This paper cites Vqa: Visual question answering.

SynRES: Towards Referring Expression Segmentation in the Wild via Synthetic Data Vqa: Visual question answering

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:47:40.174662Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:47:20.224226Z digest=sha256:676ea133509e3f985d251a09927ff04df8a517755523b19f45e8e1e10ffa8961

Observation 75d2d425-9aed-44f3-bff1-77c513afcb5e · outbound

This paper cites Coco- stuff: Thing and stuff classes in context.

SynRES: Towards Referring Expression Segmentation in the Wild via Synthetic Data Coco- stuff: Thing and stuff classes in context

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:47:39.965237Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:47:20.354744Z digest=sha256:4228cd52c419e1053b34c54b56a5121343bce1e4918795e62845ac98c702e540

Observation 0040dbe8-b07a-4196-b433-0c795bcf9153 · outbound

This paper cites Detect what you can: De- tecting and representing objects using holistic models and body parts.

SynRES: Towards Referring Expression Segmentation in the Wild via Synthetic Data Detect what you can: De- tecting and representing objects using holistic models and body parts

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:47:39.855083Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:47:20.582089Z digest=sha256:fa0b383358ab4496a0818567980c71084d6fdaf0fe14724cfb93cfc4efaf6420

Observation 29e63728-7e7c-4d1d-840a-d251c80e6c04 · outbound

This paper cites Sam4mllm: Enhance multi- modal large language model for referring expression seg- mentation.

SynRES: Towards Referring Expression Segmentation in the Wild via Synthetic Data Sam4mllm: Enhance multi- modal large language model for referring expression seg- mentation

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:47:39.662221Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:47:20.719249Z digest=sha256:c5691f8b619421a5be254226f828d23ee45c5e6a340702cebe160b1c5339fef0

Observation 593d3b04-b604-4091-b064-ecf33998db1b · outbound

This paper cites The cityscapes dataset for semantic urban scene understanding.

SynRES: Towards Referring Expression Segmentation in the Wild via Synthetic Data The cityscapes dataset for semantic urban scene understanding

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:47:39.388987Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:47:20.814327Z digest=sha256:47585274891d6bc30957cf425b175393cd5555cccce0cf339e045a205d92219e

Observation ad7a91b9-c475-43c5-a60d-ee70e36b05ab · outbound

This paper cites Divergen: Improving instance segmentation by learning wider data distribution with more diverse generative data.

SynRES: Towards Referring Expression Segmentation in the Wild via Synthetic Data Divergen: Improving instance segmentation by learning wider data distribution with more diverse generative data

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:47:39.171952Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:47:20.945150Z digest=sha256:18e9d517dc4dea41bbe2be2b19debed5fc78464f27b413a967c4baeb19b90f03

Observation 4f51af55-ece7-4008-92f8-d16910be9951 · outbound

This paper cites Finding nemo: Negative- mined mosaic augmentation for referring image segmentation.

SynRES: Towards Referring Expression Segmentation in the Wild via Synthetic Data Finding nemo: Negative- mined mosaic augmentation for referring image segmentation

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:47:38.874547Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:47:21.055373Z digest=sha256:d715a4155a8b7c725cf936b2611b4f7793ddf71b4cede66abc15ee9245d44dd0

Observation ed78cdbe-d00e-4fc0-ba65-b58e6c349938 · outbound

This paper cites Mixgen: A new multi- modal data augmentation.

SynRES: Towards Referring Expression Segmentation in the Wild via Synthetic Data Mixgen: A new multi- modal data augmentation

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:47:38.621618Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:47:21.156345Z digest=sha256:5aa9018e79b75e366bcc407479686a499cb6f65ccdd0c0798c6578708ed5fe51

Observation 6b5d6fab-f880-40aa-ab36-6cfdcd249263 · outbound

This paper cites Partimagenet: A large, high-quality dataset of parts.

SynRES: Towards Referring Expression Segmentation in the Wild via Synthetic Data Partimagenet: A large, high-quality dataset of parts

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:47:38.344668Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:47:21.270993Z digest=sha256:a7ddc1d506be7ba611e592b20f9504ab0154ee4d468010680ad869af57824e3d

Observation 4d663812-4d0e-492c-a258-9c9a1b0d5a54 · outbound

This paper cites Denoising dif- fusion probabilistic models.

SynRES: Towards Referring Expression Segmentation in the Wild via Synthetic Data Denoising dif- fusion probabilistic models

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-07T14:47:21.417306Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:47:21.417306Z digest=sha256:a12839a22b6a706aa34c3ccb0952d83b04e1ac8eaa941e6cb3954369ae81184f

Observation cc5a2dd0-2c42-4778-800d-c7a349015700 · outbound

This paper cites Beyond one-to-one: Rethinking the referring image segmentation.

SynRES: Towards Referring Expression Segmentation in the Wild via Synthetic Data Beyond one-to-one: Rethinking the referring image segmentation

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:47:38.073131Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:47:21.520435Z digest=sha256:2e28fbe61e10a2dba2e492a55a20c9c9f99d44557605bead94987ac52e5d3456

Observation 982feeb2-8e80-48da-ad7e-2911838a7c6b · outbound

This paper cites GPT-4o System Card.

SynRES: Towards Referring Expression Segmentation in the Wild via Synthetic Data GPT-4o System Card

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-07T14:47:21.632881Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:47:21.632881Z digest=sha256:0bfb1c2e09122381e6a7c4a593dc4439ac63cf4d0d78554b8ff5ab4932a5e5b1

Observation 04aa2b42-bb28-45f9-991e-09beeba581d7 · outbound

This paper cites ARMADA: Attribute-Based Multimodal Data Augmentation.

SynRES: Towards Referring Expression Segmentation in the Wild via Synthetic Data ARMADA: Attribute-Based Multimodal Data Augmentation

Reference 14

Resolution
verified exact
local_arxiv, observed 2026-08-07T14:47:27.779282Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:47:21.743339Z digest=sha256:c72f33b7c93cfe5fcb96fd5e7a9f067b15e3802518b002cd0350c5b4177bf937

Observation 81d70f02-e4d9-499e-a0d0-6180e3306afe · outbound

This paper cites Referitgame: Referring to objects in pho- tographs of natural scenes.

SynRES: Towards Referring Expression Segmentation in the Wild via Synthetic Data Referitgame: Referring to objects in pho- tographs of natural scenes

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:47:37.826592Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:47:21.838974Z digest=sha256:03ec24eeaf8169feff0f625f7e65b2181383d17e148f2d4cde02cc59af49efd9

Observation ae266e83-b34f-4ec4-97a8-e409dbdb1892 · outbound

This paper cites Segment any- thing.

SynRES: Towards Referring Expression Segmentation in the Wild via Synthetic Data Segment any- thing

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:47:37.601437Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:47:21.921379Z digest=sha256:cd52fdf9461ce8cee5127408f73f2e16f2461f861f0a8805e0439f67984edcf6

Observation 1bacd787-1def-49b8-b1b3-46cb72466998 · outbound

This paper cites Lisa: Reasoning segmentation via large language model.

SynRES: Towards Referring Expression Segmentation in the Wild via Synthetic Data Lisa: Reasoning segmentation via large language model

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:47:37.363436Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:47:22.006435Z digest=sha256:e089ee2192a618f5a1c30ca833d0d2bab7ca80edb716bf0cfb927be4ab12cc8b

Observation 698c7ee5-b400-4e3a-9262-594cd968a8ac · outbound

This paper cites Bigdatasetgan: Synthe- sizing imagenet with pixel-wise annotations.

SynRES: Towards Referring Expression Segmentation in the Wild via Synthetic Data Bigdatasetgan: Synthe- sizing imagenet with pixel-wise annotations

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:47:37.057411Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:47:22.120552Z digest=sha256:73cc9e5214fc97359f46f697c0a88df7fce5fa454e317583eaa1c4651402118f

Observation b7c1f8cc-3b97-41dd-b890-b773d04e95ec · outbound

This paper cites Gres: Gener- alized referring expression segmentation.

SynRES: Towards Referring Expression Segmentation in the Wild via Synthetic Data Gres: Gener- alized referring expression segmentation

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:47:36.829544Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:47:22.239095Z digest=sha256:9f82e44297c950e5c0107ab0f68e58d8de2037b6f690f38bd0bf7eb7ba86e932

Observation f3e6ca16-c619-4cac-8b4d-c154571f153c · outbound

This paper cites Improved baselines with visual instruction tuning.

SynRES: Towards Referring Expression Segmentation in the Wild via Synthetic Data Improved baselines with visual instruction tuning

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:47:36.591923Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:47:22.322938Z digest=sha256:67c4a93368f4fad7e7806cce1de89ccdf2af5826e4f8cc243a799148e5ba9b76

Observation e1829636-215f-4de0-8a45-31ebf64feeab · outbound

This paper cites Visual instruction tuning.

SynRES: Towards Referring Expression Segmentation in the Wild via Synthetic Data Visual instruction tuning

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:47:36.367977Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:47:22.405143Z digest=sha256:1b51251cbc14b5d72850abea99aae4ac7204530de61b5a0a36c9a43e07485954

Observation 7b0b0c59-ebbc-4ddc-bb91-8141aea95786 · outbound

This paper cites Learning multimodal data augmentation in feature space.

SynRES: Towards Referring Expression Segmentation in the Wild via Synthetic Data Learning multimodal data augmentation in feature space

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:47:36.085520Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:47:22.497171Z digest=sha256:9e4a1946d4bf7c284e609d500f8ab2ad7d45f350746adc1d328846f3e7bcd340

Observation 2b756d41-74bb-40b7-8fa0-0565df969ac0 · outbound

This paper cites Generation and com- prehension of unambiguous object descriptions.

SynRES: Towards Referring Expression Segmentation in the Wild via Synthetic Data Generation and com- prehension of unambiguous object descriptions

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:47:35.769701Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:47:22.582842Z digest=sha256:99d8108b656a371c47c7f5035ab1f4558f8b51acdd30478a8f69c55155e13262

Observation 2d0486e6-a933-42fd-a5ad-faba589ad9f0 · outbound

This paper cites Arm- bench: An object-centric benchmark dataset for robotic ma- nipulation.

SynRES: Towards Referring Expression Segmentation in the Wild via Synthetic Data Arm- bench: An object-centric benchmark dataset for robotic ma- nipulation

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:47:35.465501Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:47:22.655738Z digest=sha256:3272002c329124774de4d44430a8ec7cdb74cfe5ff4b3d862b62d505247219c4

Observation 376cf2ff-cf1e-4986-b3dc-81d7c8a7364a · outbound

This paper cites The mapillary vistas dataset for semantic understanding of street scenes.

SynRES: Towards Referring Expression Segmentation in the Wild via Synthetic Data The mapillary vistas dataset for semantic understanding of street scenes

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:47:35.253944Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:47:22.831522Z digest=sha256:47a3670e4f9d0ac64c8817d5fec51343ae71436e63c6e7f5950d38b42b53141b

Observation 8c7a92a1-16bf-4206-a287-7634dca742da · outbound

This paper cites Dataset diffusion: Diffusion-based synthetic data generation for pixel-level semantic segmentation.

SynRES: Towards Referring Expression Segmentation in the Wild via Synthetic Data Dataset diffusion: Diffusion-based synthetic data generation for pixel-level semantic segmentation

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:47:34.982424Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:47:22.906508Z digest=sha256:493b8e54523f747dbce207d21cf39e90184f2979bf9be6a71880bc0f0b30d2d6

Observation 6e601524-4404-4457-948b-7ebddc73b350 · outbound

This paper cites Learning transferable visual models from natural language supervision.

SynRES: Towards Referring Expression Segmentation in the Wild via Synthetic Data Learning transferable visual models from natural language supervision

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:47:34.697135Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:47:23.020067Z digest=sha256:7b8cbf52a4b172b06869f15023019d703bdf76b3d3526eb953a3f5f743c7abfa

Observation e2e5d3db-c511-4199-8a54-e25671517fc1 · outbound

This paper cites Paco: Parts and attributes of common objects.

SynRES: Towards Referring Expression Segmentation in the Wild via Synthetic Data Paco: Parts and attributes of common objects

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:47:34.330678Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:47:23.143071Z digest=sha256:5a118ecd731249d17871d3b95d62fecb0ba2464d1e8be9c1f5e2e41ad811db5e

Observation fa6d6adf-7a56-43be-8858-94a792959ffd · outbound

This paper cites Glamm: Pixel grounding large multimodal model.

SynRES: Towards Referring Expression Segmentation in the Wild via Synthetic Data Glamm: Pixel grounding large multimodal model

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:47:34.086363Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:47:23.247935Z digest=sha256:0147f63a47c77afeef0f956d89a367110f1d57fa2239cb8b3d5e68120692000a

Observation 86e7719c-f194-4a07-ab95-adc4a619e2ed · outbound

This paper cites SAM 2: Segment anything in images and videos.

SynRES: Towards Referring Expression Segmentation in the Wild via Synthetic Data SAM 2: Segment anything in images and videos

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:47:33.780743Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:47:23.370944Z digest=sha256:52541f4ba5fa1b053038ed2e8cb9c6df5a9548e43d91237c0088a1a8445a9802

Observation c4b4b0f9-a4a7-4e5d-bbd5-950cb452100a · outbound

This paper cites Pixellm: Pixel reasoning with large multimodal model.

SynRES: Towards Referring Expression Segmentation in the Wild via Synthetic Data Pixellm: Pixel reasoning with large multimodal model

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:47:33.485800Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:47:23.476000Z digest=sha256:df972e6234326228775d7db278f730cbb72743eb8ec5200dd7602f27e17b93a0

Observation 9cb64232-9601-4aaa-8e40-c94bc1fd775a · outbound

This paper cites Grounding of textual phrases in images by reconstruction.

SynRES: Towards Referring Expression Segmentation in the Wild via Synthetic Data Grounding of textual phrases in images by reconstruction

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:47:33.238447Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:47:23.585827Z digest=sha256:d0d07af5e00fb3b02b6f9509699c39075648abf452d4aad660abc13f6382080c

Observation a7751b07-0325-4461-8730-f2f0486faf37 · outbound

This paper cites CrowdHuman: A Benchmark for Detecting Human in a Crowd.

SynRES: Towards Referring Expression Segmentation in the Wild via Synthetic Data CrowdHuman: A Benchmark for Detecting Human in a Crowd

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-07T14:47:23.699161Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:47:23.699161Z digest=sha256:0eba5b8b20c03b1a471f624f4d058978abe6381a96aeb9facd65f454b653197d

Observation a09e70b1-3ba4-4121-a076-ae22b72a40aa · outbound

This paper cites Denoising diffusion implicit models.

SynRES: Towards Referring Expression Segmentation in the Wild via Synthetic Data Denoising diffusion implicit models

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:47:33.015018Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:47:23.808606Z digest=sha256:5d8f88881207dd1c539a09b615f3aff6db01be48082264f98606b9c395a143d9

Observation 5f946cb7-fbce-46cf-8317-c2e35efda692 · outbound

This paper cites DreamMask: Boosting Open-vocabulary Panoptic Segmentation with Synthetic Data.

SynRES: Towards Referring Expression Segmentation in the Wild via Synthetic Data DreamMask: Boosting Open-vocabulary Panoptic Segmentation with Synthetic Data

Reference 35

Resolution
verified exact
local_arxiv, observed 2026-08-07T14:47:27.583189Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:47:23.916439Z digest=sha256:0e3dbcbeb559eba658187f8f242107828c5cc6c059d4b62c14ad6281ce0fcff4

Observation 110b8971-dade-48ef-83e0-968aa9e3a808 · outbound

This paper cites Cris: Clip-driven referring image segmentation.

SynRES: Towards Referring Expression Segmentation in the Wild via Synthetic Data Cris: Clip-driven referring image segmentation

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:47:32.788625Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:47:24.069836Z digest=sha256:1695bd98d86b6341d61d5986d3537e3bd5cac9bf736222656496983b6774b6ad

Observation af2dc787-bdba-49e3-aa9e-79ca09ac039e · outbound

This paper cites Towards reporting bias in visual-language datasets: bimodal augmentation by decoupling object-attribute association.

SynRES: Towards Referring Expression Segmentation in the Wild via Synthetic Data Towards reporting bias in visual-language datasets: bimodal augmentation by decoupling object-attribute association

Reference 37

Resolution
verified exact
local_arxiv, observed 2026-08-07T14:47:27.366597Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:47:24.223505Z digest=sha256:2164240812d3a3a834d1fc64f0cc04b87e070b6a1aa3bd8a5bae782e47f2c9f4

Observation 311c22a3-2dc1-427d-811d-a8c2a2111c5b · outbound

This paper cites Diffumask: Synthesizing images with pixel-level annotations for semantic segmentation using diffu- sion models.

SynRES: Towards Referring Expression Segmentation in the Wild via Synthetic Data Diffumask: Synthesizing images with pixel-level annotations for semantic segmentation using diffu- sion models

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:47:32.549960Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:47:24.377245Z digest=sha256:5bf9ba6b42bba7618e1a37417305a8f80a6cbb2bd20689d6ee126841f44bb4c4

Observation 65cc8935-735d-46fc-a756-96c111073256 · outbound

This paper cites Gsva: Generalized segmentation via multimodal large language models.

SynRES: Towards Referring Expression Segmentation in the Wild via Synthetic Data Gsva: Generalized segmentation via multimodal large language models

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:47:32.290075Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:47:24.527351Z digest=sha256:5aac4c7bb74665fd1bf4f7133c5cb158288fc6888bd6c44a4441dc8ead925dd7

Observation 664af2dd-7586-4e2f-b56f-a4755bf3bebb · outbound

This paper cites SANA: Efficient High-Resolution Image Synthesis with Linear Diffusion Transformers.

SynRES: Towards Referring Expression Segmentation in the Wild via Synthetic Data SANA: Efficient High-Resolution Image Synthesis with Linear Diffusion Transformers

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-07T14:47:24.674218Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:47:24.674218Z digest=sha256:df214dd3092c8f965243f5d94232bd648c43e587f87c8ef2957ed6ea1d1d4666

Observation f9c2fbe1-7532-40b5-a8b8-a64bd7d7408e · outbound

This paper cites Mosaicfusion: Diffusion models as data augmenters for large vocabulary instance segmentation.

SynRES: Towards Referring Expression Segmentation in the Wild via Synthetic Data Mosaicfusion: Diffusion models as data augmenters for large vocabulary instance segmentation

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:47:32.060147Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:47:24.804534Z digest=sha256:95f8038f91a72007caff468cd4c6c29244a6be8cf27907e957ac9fe37537850d

Observation 896b32b1-bbf4-4502-85ab-79dc11b0e4f7 · outbound

This paper cites Bridging vision and language encoders: Parameter-efficient tuning for referring image segmentation.

SynRES: Towards Referring Expression Segmentation in the Wild via Synthetic Data Bridging vision and language encoders: Parameter-efficient tuning for referring image segmentation

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:47:31.877017Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:47:24.935045Z digest=sha256:99a3f4bbb42af85cd1705584f9c5a721dbf9da87d57bbcbbbf93e558bf8a6664

Observation d107d3a5-db7d-4cc3-b2c0-5f562362d384 · outbound

This paper cites Panoptic scene graph gen- eration.

SynRES: Towards Referring Expression Segmentation in the Wild via Synthetic Data Panoptic scene graph gen- eration

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:47:31.620806Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:47:25.078638Z digest=sha256:8188038c098a4438821832dc570acd9d5d98e12db43973a952b59252cd6d3927

Observation 1365327d-86c0-4d4a-b647-c82600da3a52 · outbound

This paper cites Freemask: Synthetic images with dense annotations make stronger segmentation models.

SynRES: Towards Referring Expression Segmentation in the Wild via Synthetic Data Freemask: Synthetic images with dense annotations make stronger segmentation models

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:47:31.349947Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:47:25.211282Z digest=sha256:36b97fa06e56f8279e622ed98d7325db0dc2b71b54cd30af99dec54d5435aeff

Observation b8f52a39-dab2-41d0-8719-3e2d99ac9dbc · outbound

This paper cites Lavt: Language-aware vi- sion transformer for referring image segmentation.

SynRES: Towards Referring Expression Segmentation in the Wild via Synthetic Data Lavt: Language-aware vi- sion transformer for referring image segmentation

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:47:31.049158Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:47:25.303573Z digest=sha256:cde26e6c4f6557465f958bf97d45f68ea02b35f386ea0dfbd32c0d15a9a1e4fa

Observation 5b90db18-c264-45a0-a000-dc216eddebb7 · outbound

This paper cites Seggen: Supercharging segmentation models with text2mask and mask2img synthesis.

SynRES: Towards Referring Expression Segmentation in the Wild via Synthetic Data Seggen: Supercharging segmentation models with text2mask and mask2img synthesis

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:47:30.759804Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:47:25.460160Z digest=sha256:5e5d89733232abbc3b770c6ec2170e14a87fb81d927c3022cd024a1a4bc2f14b

Observation 9e21ba7b-94d9-4cd1-8b4c-13e978fbf85c · outbound

This paper cites Pix2Cap-COCO: Advancing Visual Comprehension via Pixel-Level Captioning.

SynRES: Towards Referring Expression Segmentation in the Wild via Synthetic Data Pix2Cap-COCO: Advancing Visual Comprehension via Pixel-Level Captioning

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-07T14:47:25.624443Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:47:25.624443Z digest=sha256:45a1fe0f16a1b73fa11ae523dafbb65982b1350d4fc3c4a942641d21e29e341c

Observation be04d225-2071-4d7c-9095-03a4d872127f · outbound

This paper cites From image descriptions to visual denotations: New similarity metrics for semantic inference over event descrip- tions.

SynRES: Towards Referring Expression Segmentation in the Wild via Synthetic Data From image descriptions to visual denotations: New similarity metrics for semantic inference over event descrip- tions

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:47:30.484684Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:47:25.738988Z digest=sha256:cdb02777c388390694fa8918efc039145d17ca978ff94316c24b12e7ff774748

Observation acee422e-2b11-4869-b89b-d5faaecd52b9 · outbound

This paper cites Coca: Contrastive captioners are image-text foundation models.

SynRES: Towards Referring Expression Segmentation in the Wild via Synthetic Data Coca: Contrastive captioners are image-text foundation models

Reference 49

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:47:30.246850Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:47:25.851673Z digest=sha256:a914992e4aa405661d8822809dfb7e75c64f9a614c08751a04b7ffe7d7c0e8f1

Observation 298725ce-b1b7-465e-9404-4ce1c629a6b5 · outbound

This paper cites Modeling context in referring expressions.

SynRES: Towards Referring Expression Segmentation in the Wild via Synthetic Data Modeling context in referring expressions

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:47:30.004079Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:47:25.960341Z digest=sha256:c667464157bc11751ad9841f7b1c166aabd4065ddf868e7d5caa85e153d33965

Observation 53ee01ac-5595-4c72-ad2d-3706a951b1f9 · outbound

This paper cites Pseudo- ris: Distinctive pseudo-supervision generation for referring image segmentation.

SynRES: Towards Referring Expression Segmentation in the Wild via Synthetic Data Pseudo- ris: Distinctive pseudo-supervision generation for referring image segmentation

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:47:29.677320Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:47:26.080046Z digest=sha256:25abfc1c3201dabc367340f22ee4544b1a6cdf495a53c5ba1682ce7d079533bd

Observation 029d00c9-a3b2-46ca-a571-56dae9d18660 · outbound

This paper cites Revisiting counterfactual prob- lems in referring expression comprehension.

SynRES: Towards Referring Expression Segmentation in the Wild via Synthetic Data Revisiting counterfactual prob- lems in referring expression comprehension

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:47:29.418166Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:47:26.169375Z digest=sha256:585f23ad2343dc4ecbbda1b951eb1834a6ec12ec8b0df29f05cc3318be960a36

Observation f413083f-ea99-4897-bd65-4ed33e1fddfe · outbound

This paper cites Datasetgan: Efficient labeled data factory with minimal human effort.

SynRES: Towards Referring Expression Segmentation in the Wild via Synthetic Data Datasetgan: Efficient labeled data factory with minimal human effort

Reference 53

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:47:29.165630Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:47:26.285795Z digest=sha256:6fd1f536941c8283ec5b098a5d19d0023a8b31e602d3488ed3fe468e876de3ac

Observation eea956fa-8b55-47b8-904c-11aabd935bef · outbound

This paper cites EVF-SAM: Early Vision-Language Fusion for Text-Prompted Segment Anything Model.

SynRES: Towards Referring Expression Segmentation in the Wild via Synthetic Data EVF-SAM: Early Vision-Language Fusion for Text-Prompted Segment Anything Model

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-07T14:47:26.416689Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:47:26.416689Z digest=sha256:f51bfa6207d70a30cddf50c62de9e596b38fff251b8e394b6c450a5f28f835aa

Observation ca4ff209-3265-47d3-8c01-6989a655964e · outbound

This paper cites Psalm: Pixelwise segmentation with large multi-modal model.

SynRES: Towards Referring Expression Segmentation in the Wild via Synthetic Data Psalm: Pixelwise segmentation with large multi-modal model

Reference 55

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:47:28.928964Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:47:26.551994Z digest=sha256:c90bd6e3d681301327358520b302c3d6f559cb476c7bf95bc913c03163292cb5

Observation f7082026-268a-4aff-b4bf-f8b162d5061d · outbound

This paper cites X-paste: Revisiting scalable copy-paste for in- stance segmentation using clip and stablediffusion.

SynRES: Towards Referring Expression Segmentation in the Wild via Synthetic Data X-paste: Revisiting scalable copy-paste for in- stance segmentation using clip and stablediffusion

Reference 56

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:47:28.685348Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:47:26.658210Z digest=sha256:d1a248dd18edbe58d3aeccf76c4840c1999e06acb22cc2240cec890abb2592d1

Observation faaab8c1-8e33-4b9d-bde0-3cda016368a8 · outbound

This paper cites Unleashing text-to-image diffusion models for visual perception.

SynRES: Towards Referring Expression Segmentation in the Wild via Synthetic Data Unleashing text-to-image diffusion models for visual perception

Reference 57

Resolution
unresolved
no resolver link, observed 2026-08-07T14:47:26.760787Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:47:26.760787Z digest=sha256:c80e31b1345887caf010cc0865fffeb23405511463a519cf2bfa540b0b27b455

Observation 7a5485ab-0c7b-4bbc-a4f3-e759d859ec47 · outbound

This paper cites Scene parsing through ade20k dataset.

SynRES: Towards Referring Expression Segmentation in the Wild via Synthetic Data Scene parsing through ade20k dataset

Reference 58

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:47:28.461809Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:47:26.890631Z digest=sha256:613e0b24e047da7b5c52fafcd1b6a9f2c04328f6e479231327fff537539ce631

Observation a9a2d513-693c-4783-b193-ac6be2c89da3 · outbound

This paper cites Generalized decoding for pixel, image, and language.

SynRES: Towards Referring Expression Segmentation in the Wild via Synthetic Data Generalized decoding for pixel, image, and language

Reference 59

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:47:28.237992Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:47:27.016843Z digest=sha256:ac458f8e0de8fe2bdbf3d9a4090de808a1ff334eeb0374cffc5d2e057b339097

Observation 926b136e-3d50-499b-99db-ad42737ca299 · outbound

This paper cites the cat sitting on the bench next to big green wooden boat in the center of the image.

SynRES: Towards Referring Expression Segmentation in the Wild via Synthetic Data the cat sitting on the bench next to big green wooden boat in the center of the image

Reference 60

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:47:27.953849Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:47:27.120881Z digest=sha256:8457dc48566736a80bd667567d607f6ab3c52c54d3008d716d3afbd7c5c47f50

Pith citing papers

No inbound Pith citation observations are available.