Pith. sign in

Paper Citation Record · LEDGER

Stepping Out of Similar Semantic Space for Open-Vocabulary Segmentation

As of 7 August 2026, this Paper Citation Record lists 66 of 66 outbound references and 0 inbound Pith citation observations for arXiv:2506.16058.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.16058 v2

Coverage vector

measured 66 of 66 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T23:50:23.565560Z

measured 66 of 66 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

66 of 66 outbound references displayed

  • verified exact3
  • verified fuzzy43
  • unresolved20
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation a97559ff-9004-4296-ab05-ec06c700efd3 · outbound

This paper cites Self-calibrated clip for training-free open-vocabulary segmentation.

Stepping Out of Similar Semantic Space for Open-Vocabulary Segmentation Self-calibrated clip for training-free open-vocabulary segmentation

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-06T23:50:23.250219Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:50:23.250219Z digest=sha256:ff8dac0b7ed2b5f8d47e4052ea8211b899aa29fa97d5aaa4199e13e90e19350c

Observation 8ddce7f4-5588-41a2-83f5-52828cd4f189 · outbound

This paper cites UniVG-R1: Reasoning Guided Universal Visual Grounding with Reinforcement Learning.

Stepping Out of Similar Semantic Space for Open-Vocabulary Segmentation UniVG-R1: Reasoning Guided Universal Visual Grounding with Reinforcement Learning

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-06T23:50:23.255593Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:50:23.255593Z digest=sha256:9b71c33256ee0017d1c3749f883b8f7a3a6ca3ac6616bc3cbe3c350d54f70300

Observation 97e86502-505d-4e63-a351-faf64d93292e · outbound

This paper cites Brostow, Jamie Shotton, Julien Fauqueur, and Roberto Cipolla.

Stepping Out of Similar Semantic Space for Open-Vocabulary Segmentation Brostow, Jamie Shotton, Julien Fauqueur, and Roberto Cipolla

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:50:24.726791Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T23:50:23.261190Z digest=sha256:97956aeda5f22942f0e4cef495ea9e083e7c36d5a2fd51f91aa3569af59ab2bf

Observation 668c0a39-61f2-4d63-b2c7-837a4f063b19 · outbound

This paper cites Zero-shot semantic segmentation.

Stepping Out of Similar Semantic Space for Open-Vocabulary Segmentation Zero-shot semantic segmentation

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:50:24.711443Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T23:50:23.266818Z digest=sha256:50021f4dbeccc0386ab1009319a30527adf7da372d7e1cbc21ad19a411de0122

Observation c50bfe99-7026-4e63-8648-f42be4d18a4a · outbound

This paper cites End-to- end object detection with transformers.

Stepping Out of Similar Semantic Space for Open-Vocabulary Segmentation End-to- end object detection with transformers

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:50:24.696514Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T23:50:23.271584Z digest=sha256:d22eae7b02c4b5f06fc8a7ea1647913fb5cff80d5acd20d23b47585a7e82bf2f

Observation e38a447f-63b6-4c93-95d4-88c76ddd1afe · outbound

This paper cites an unresolved cited work.

Stepping Out of Similar Semantic Space for Open-Vocabulary Segmentation Unresolved cited work

Reference 6

Resolution
unresolved
raw_fallback, observed 2026-08-06T23:50:24.681255Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T23:50:23.276968Z digest=sha256:2b17894b4ed9f39ea6ca821eef125ae9357a769bee214284f71792202aadba00

Observation 97cb7be6-a06c-463a-8359-e2eb0820f02c · outbound

This paper cites UNITER: UNiversal Image-TExt Representation Learning.

Stepping Out of Similar Semantic Space for Open-Vocabulary Segmentation UNITER: UNiversal Image-TExt Representation Learning

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-06T23:50:23.281943Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:50:23.281943Z digest=sha256:1bb12fb6a997a0d1964a0356b450e32f11eb446fe97c3e47ad94362a74e300f4

Observation cab951f1-9f45-4213-978b-f68eea78264e · outbound

This paper cites Mask2Former for Video Instance Segmentation.

Stepping Out of Similar Semantic Space for Open-Vocabulary Segmentation Mask2Former for Video Instance Segmentation

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-06T23:50:23.286782Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:50:23.286782Z digest=sha256:ec98596c4433d990290da6ba96b47297ba8a615ef8ae07745db50738c8c74bf5

Observation c5b765f5-6d38-48ca-89e7-f31e3eecfc23 · outbound

This paper cites Schwing, Alexan- der Kirillov, and Rohit Girdhar.

Stepping Out of Similar Semantic Space for Open-Vocabulary Segmentation Schwing, Alexan- der Kirillov, and Rohit Girdhar

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-06T23:50:23.292230Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:50:23.292230Z digest=sha256:8325c08e4971a9c55256dae49ef28e6407936b489e184f25fc997b9a3e6a8ae9

Observation d3f43add-dd73-459f-9379-39f466b93474 · outbound

This paper cites Cat-seg: Cost aggregation for open-vocabulary semantic segmenta- tion.

Stepping Out of Similar Semantic Space for Open-Vocabulary Segmentation Cat-seg: Cost aggregation for open-vocabulary semantic segmenta- tion

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:50:24.656705Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T23:50:23.297117Z digest=sha256:bbfe9e60cfcdadb572e27b26555be1213df6a92c1b908071d10cf91a58067465

Observation 4723c5f7-8b26-4713-a00e-80203d198265 · outbound

This paper cites The cityscapes dataset for semantic urban scene understanding.

Stepping Out of Similar Semantic Space for Open-Vocabulary Segmentation The cityscapes dataset for semantic urban scene understanding

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-06T23:50:23.301829Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:50:23.301829Z digest=sha256:7ede5d97ba10219836c5d3ce3874c5da0ce7a3becf6e7b47b67c92b9a0e5001c

Observation e4678ec7-af9e-4dda-ac7b-b7463e10667d · outbound

This paper cites De- coupling zero-shot semantic segmentation.

Stepping Out of Similar Semantic Space for Open-Vocabulary Segmentation De- coupling zero-shot semantic segmentation

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:50:24.630931Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T23:50:23.307089Z digest=sha256:c98dfdc71552b7b2c133a58fac004f1e893608b5f19092ab0ee7f06156f2c843

Observation 9f2bc313-d6c4-40a1-8738-f963b9a62445 · outbound

This paper cites The pascal visual object classes challenge: A retrospective.IJCV, 111:98–136, 2015.

Stepping Out of Similar Semantic Space for Open-Vocabulary Segmentation The pascal visual object classes challenge: A retrospective.IJCV, 111:98–136, 2015

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:50:24.615216Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T23:50:23.311413Z digest=sha256:ec486c42234353962ad006955d45ca9d96f8794f578ce3bfb10248c3c21cd851

Observation 0d1d0ab4-f453-41f8-a227-bfef21cb9f24 · outbound

This paper cites Large-scale unsu- pervised semantic segmentation.

Stepping Out of Similar Semantic Space for Open-Vocabulary Segmentation Large-scale unsu- pervised semantic segmentation

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:50:24.600200Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T23:50:23.316014Z digest=sha256:5cdc35cded180a46d68823dee59839cec039e2d815467fb8ab527ce33d456e3f

Observation edc5a565-09b0-4bc9-8784-dc05196a8f99 · outbound

This paper cites Scaling Open-Vocabulary Image Segmentation with Image-Level Labels.

Stepping Out of Similar Semantic Space for Open-Vocabulary Segmentation Scaling Open-Vocabulary Image Segmentation with Image-Level Labels

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-06T23:50:23.320589Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:50:23.320589Z digest=sha256:4661d102b531f793d1d4ef3166793012f49adaaeae5589eb5779b430d3d19c21

Observation 9e39c27a-a7ff-4470-9077-2fd5cb6a8031 · outbound

This paper cites Random walks for image segmentation.

Stepping Out of Similar Semantic Space for Open-Vocabulary Segmentation Random walks for image segmentation

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:50:24.584747Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T23:50:23.325420Z digest=sha256:ba90e49412c87f3ac0a704ade6a055f7760935e24a8b4684572b85c9feea4854

Observation 39bb6abf-29bb-4c4f-8712-ea9fb4931231 · outbound

This paper cites Global knowledge calibration for fast open-vocabulary segmentation.

Stepping Out of Similar Semantic Space for Open-Vocabulary Segmentation Global knowledge calibration for fast open-vocabulary segmentation

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:50:24.567916Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T23:50:23.330380Z digest=sha256:7dd38e7b0c11533054eb0f0768a3cf559d811e7be31e8ac2c77b1b5604b5099d

Observation 25db6e78-78b1-40f8-b904-20340d80d6c3 · outbound

This paper cites Primitive gener- ation and semantic-related alignment for universal zero-shot segmentation.

Stepping Out of Similar Semantic Space for Open-Vocabulary Segmentation Primitive gener- ation and semantic-related alignment for universal zero-shot segmentation

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:50:24.552376Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T23:50:23.335212Z digest=sha256:63522cef9815d2fd033485f1ede0c3d2c1935d03c21a5a9ddb56ac172828767e

Observation b676ed18-8618-4d5d-a61e-2ef7499ba897 · outbound

This paper cites Planning-oriented autonomous driving.

Stepping Out of Similar Semantic Space for Open-Vocabulary Segmentation Planning-oriented autonomous driving

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-06T23:50:23.340138Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:50:23.340138Z digest=sha256:b27d558fa09e5bda773a9ff090e1c8266b90b9cf1d29e8de1fc947493585d54a

Observation 3268e0aa-7e79-441c-93b2-a1f7b2c8fd6c · outbound

This paper cites Densely Connected Parameter-Efficient Tuning for Referring Image Segmentation.

Stepping Out of Similar Semantic Space for Open-Vocabulary Segmentation Densely Connected Parameter-Efficient Tuning for Referring Image Segmentation

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-06T23:50:23.344365Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:50:23.344365Z digest=sha256:536269d08101091fca1106216b12d5674f7f7c69b8e4438b9c745ae274a1552b

Observation de7f347b-2b56-4e45-98e4-5c4184e9f8af · outbound

This paper cites Segment and Caption Anything.

Stepping Out of Similar Semantic Space for Open-Vocabulary Segmentation Segment and Caption Anything

Reference 21

Resolution
verified exact
local_arxiv, observed 2026-08-06T23:50:23.795324Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T23:50:23.348944Z digest=sha256:01b18e592ce2e78059f502c1b949332214686fca534a3a178dcbaf1ac144909b

Observation cc55fbdc-facc-4b62-86d5-2b7d9f022012 · outbound

This paper cites Proxydet: Synthesizing proxy novel classes via classwise mixup for open-vocabulary object detection.

Stepping Out of Similar Semantic Space for Open-Vocabulary Segmentation Proxydet: Synthesizing proxy novel classes via classwise mixup for open-vocabulary object detection

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:50:24.525350Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T23:50:23.354116Z digest=sha256:258d8fb06f9c225c89b929980002183e2a0b65d18f19087bd71706c49c067d0a

Observation 98fa837c-4a11-4694-8f06-5a56b65b9df6 · outbound

This paper cites Le, Yun-Hsuan Sung, Zhen Li, and Tom Duerig.

Stepping Out of Similar Semantic Space for Open-Vocabulary Segmentation Le, Yun-Hsuan Sung, Zhen Li, and Tom Duerig

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:50:24.511326Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T23:50:23.359615Z digest=sha256:ff59ff6611d1e25fad1b10ac4c15c62504fc6da2161ae6fc223e22d246d52366

Observation d147be5f-999c-4b9c-a8d1-e16f67e5a740 · outbound

This paper cites Scaling up visual and vision-language representation learning with noisy text supervision.

Stepping Out of Similar Semantic Space for Open-Vocabulary Segmentation Scaling up visual and vision-language representation learning with noisy text supervision

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:50:24.495762Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T23:50:23.364378Z digest=sha256:e52806ba9bc763c0bb16614d370d0fbf7f685dd7c553461e64823f91ab20b580

Observation 8e05eb0c-bd88-40b9-a4bf-6d017e7c101b · outbound

This paper cites Learning Mask-aware CLIP Representations for Zero-Shot Segmentation.

Stepping Out of Similar Semantic Space for Open-Vocabulary Segmentation Learning Mask-aware CLIP Representations for Zero-Shot Segmentation

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-06T23:50:23.369382Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:50:23.369382Z digest=sha256:16567e05d71ef52d70ee17a0bed8f1b734993a08ebbdc3fdfdf183261a0ed845

Observation 1fc46008-179d-4f55-97b9-d16c91fbdb51 · outbound

This paper cites Collaborative vision-text rep- resentation optimizing for open-vocabulary segmentation.

Stepping Out of Similar Semantic Space for Open-Vocabulary Segmentation Collaborative vision-text rep- resentation optimizing for open-vocabulary segmentation

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:50:24.479936Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T23:50:23.374458Z digest=sha256:7eb946a8d3cd6604f45976357adc4f6047ed4a901b935e7fec1f9ddb2ccd676b

Observation 498432c4-95c4-45e0-855e-1e3a1fe63fd2 · outbound

This paper cites Weinberger, Serge J.

Stepping Out of Similar Semantic Space for Open-Vocabulary Segmentation Weinberger, Serge J

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:50:24.464721Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T23:50:23.379399Z digest=sha256:fb4fa8931a3bf32ca567d8887440468a037d98487a069d803524c4e6d406e4f0

Observation 0a889458-a699-4ddb-8b2b-e8528c40033f · outbound

This paper cites Unicoder-vl: A universal encoder for vision and lan- guage by cross-modal pre-training.

Stepping Out of Similar Semantic Space for Open-Vocabulary Segmentation Unicoder-vl: A universal encoder for vision and lan- guage by cross-modal pre-training

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:50:24.448213Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T23:50:23.384236Z digest=sha256:d270b78d31bc8b2bc5528b12c6ccb7b3b9726985230327ee8795aee0feffb25f

Observation 0ebd4417-b1aa-4019-927d-55a82dc79e9f · outbound

This paper cites Ordinalclip: Learning rank prompts for language-guided ordinal regression.

Stepping Out of Similar Semantic Space for Open-Vocabulary Segmentation Ordinalclip: Learning rank prompts for language-guided ordinal regression

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:50:24.432092Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T23:50:23.388956Z digest=sha256:5bcb0c97ec65879ec898e254958e40f42ef954ebde8a6cf6698fca0010537427

Observation c811938a-aac7-4ac5-8e09-acf2ef80452c · outbound

This paper cites Oscar: Object-semantics aligned pre-training for vision-language tasks.

Stepping Out of Similar Semantic Space for Open-Vocabulary Segmentation Oscar: Object-semantics aligned pre-training for vision-language tasks

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:50:24.416066Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T23:50:23.393647Z digest=sha256:763d14547a89b172e7cc0f3b1622f676286f908fa4558940048f2529c3e53abb

Observation 68a2fe49-d63e-45c2-bf79-9736b651e201 · outbound

This paper cites Open-Vocabulary Semantic Segmentation with Mask-adapted CLIP.

Stepping Out of Similar Semantic Space for Open-Vocabulary Segmentation Open-Vocabulary Semantic Segmentation with Mask-adapted CLIP

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-06T23:50:23.398271Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:50:23.398271Z digest=sha256:445f8613af527891f56fee4338b7995a39d2f6e62ff601689ae830c614a6f152

Observation fd7982dc-77c4-4d1e-ab7a-7ab4b057b9c1 · outbound

This paper cites Belongie, James Hays, Pietro Perona, Deva Ramanan, Piotr Doll ´ar, and C.

Stepping Out of Similar Semantic Space for Open-Vocabulary Segmentation Belongie, James Hays, Pietro Perona, Deva Ramanan, Piotr Doll ´ar, and C

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:50:24.401125Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T23:50:23.402939Z digest=sha256:aa7055f64f13ca2bbc04ac250f2e33435bf0475055781c61e0c6f0e64416bfe1

Observation 6efe546e-7a4c-4061-8f09-b8ab3b5e83b7 · outbound

This paper cites Quality- aware and selective prior enhancement memory network for video object segmentation.

Stepping Out of Similar Semantic Space for Open-Vocabulary Segmentation Quality- aware and selective prior enhancement memory network for video object segmentation

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:50:24.385247Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T23:50:23.407360Z digest=sha256:3ff8f1b9f222d4003e32d2714d7d7dabba1055528d4a0eac8661d0bec0aed317

Observation 66b8b1e0-6604-4c24-998d-108842b10359 · outbound

This paper cites Global spectral filter memory network for video object segmentation.

Stepping Out of Similar Semantic Space for Open-Vocabulary Segmentation Global spectral filter memory network for video object segmentation

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:50:24.370046Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T23:50:23.411715Z digest=sha256:055ef4d2ad1643cff7fa8a8d9aca90a1d52dfb3a371f6eeeb6a736d618236959

Observation 05c253fb-20d2-474e-b290-67a9a5a52fb9 · outbound

This paper cites Learning quality-aware dynamic memory for video object segmentation.

Stepping Out of Similar Semantic Space for Open-Vocabulary Segmentation Learning quality-aware dynamic memory for video object segmentation

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:50:24.355059Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T23:50:23.416391Z digest=sha256:a2c0552adf4b2be31aedf0b812dc16cf8d94c61058c92b8bd16705cdabce2ad9

Observation 4568aa7d-2900-49b7-a87c-311f3614fb37 · outbound

This paper cites Universal Segmentation at Arbitrary Granularity with Language Instruction.

Stepping Out of Similar Semantic Space for Open-Vocabulary Segmentation Universal Segmentation at Arbitrary Granularity with Language Instruction

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-06T23:50:23.420753Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:50:23.420753Z digest=sha256:797163bd2890c78f6f844badcd5db0bb85c4a312c44eb854d61387e57dbac775

Observation baedb819-8184-4044-9af5-e317f128b7de · outbound

This paper cites Open-vocabulary segmentation with semantic-assisted calibration.

Stepping Out of Similar Semantic Space for Open-Vocabulary Segmentation Open-vocabulary segmentation with semantic-assisted calibration

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:50:24.340256Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T23:50:23.426153Z digest=sha256:a669056a60005875bcd1a9d9aa37a017a7a64495d369fec43ddf6896e5fe4e33

Observation 0b47705e-28cb-468b-a303-fe33b669b1b1 · outbound

This paper cites Learning high-quality dynamic memory for video object segmentation.

Stepping Out of Similar Semantic Space for Open-Vocabulary Segmentation Learning high-quality dynamic memory for video object segmentation

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:50:24.324665Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T23:50:23.430987Z digest=sha256:771625bbf736459c89e9300303a2a3d9e74d9ce3caec6c9ce4bbcb0b070f613e

Observation 4358f59c-72f4-4cdf-905f-4afcf8fea044 · outbound

This paper cites ThinkBot: Embodied Instruction Following with Thought Chain Reasoning.

Stepping Out of Similar Semantic Space for Open-Vocabulary Segmentation ThinkBot: Embodied Instruction Following with Thought Chain Reasoning

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-06T23:50:23.435750Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:50:23.435750Z digest=sha256:72950e0530b49c05aabbb65af37f0c9cc915d367ab8add1eb8b4e1a4c79d79db

Observation 1b876d93-17fc-45af-9517-fa78d3341338 · outbound

This paper cites Vilbert: Pretraining task-agnostic visiolinguistic representations for vision-and-language tasks.

Stepping Out of Similar Semantic Space for Open-Vocabulary Segmentation Vilbert: Pretraining task-agnostic visiolinguistic representations for vision-and-language tasks

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:50:24.310125Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T23:50:23.440778Z digest=sha256:7e7a808265f2db09fd64045a5bd50828d2c372696b941a95566e67c3880af605

Observation 08c7c79b-2c56-4764-b50f-6073c683c43e · outbound

This paper cites SOC: Semantic-Assisted Object Cluster for Referring Video Object Segmentation.

Stepping Out of Similar Semantic Space for Open-Vocabulary Segmentation SOC: Semantic-Assisted Object Cluster for Referring Video Object Segmentation

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-06T23:50:23.445973Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:50:23.445973Z digest=sha256:f99772760d1f894b7dbd824dad4b7c2bb0dcf130dfbafe574afb46338eb2b8e9

Observation 3ba8fd79-3fbd-4731-bda1-b25a15a2fbd1 · outbound

This paper cites CoHD: A Counting-Aware Hierarchical Decoding Framework for Generalized Referring Expression Segmentation.

Stepping Out of Similar Semantic Space for Open-Vocabulary Segmentation CoHD: A Counting-Aware Hierarchical Decoding Framework for Generalized Referring Expression Segmentation

Reference 42

Resolution
verified exact
local_arxiv, observed 2026-08-06T23:50:23.696452Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T23:50:23.450624Z digest=sha256:e872197a010b9f9f4e39ce36ced26eb58ed0eea9e87b14442b261339575a42cc

Observation 881002cc-da08-4e94-9b36-86b82af250f6 · outbound

This paper cites Matrix analysis and applied linear algebra.

Stepping Out of Similar Semantic Space for Open-Vocabulary Segmentation Matrix analysis and applied linear algebra

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:50:24.294036Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T23:50:23.456174Z digest=sha256:3f825e5c6e3556ee717ac6481585d512f189dc82a14aed340993adc563f71db6

Observation 92649031-db64-4c2c-9ab0-735192d697d6 · outbound

This paper cites an unresolved cited work.

Stepping Out of Similar Semantic Space for Open-Vocabulary Segmentation Unresolved cited work

Reference 44

Resolution
unresolved
raw_fallback, observed 2026-08-06T23:50:24.279433Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T23:50:23.460868Z digest=sha256:081db866e5ca1feb74d8183ccf0d13ac18d281a38c666e9500d8b6deffd51b20

Observation 72488162-0515-4ac4-886e-9cf16112dca5 · outbound

This paper cites Siri: A simple selective retraining mechanism for transformer-based visual grounding.

Stepping Out of Similar Semantic Space for Open-Vocabulary Segmentation Siri: A simple selective retraining mechanism for transformer-based visual grounding

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:50:24.264057Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T23:50:23.465327Z digest=sha256:345bf39ae1cfad6527a0dbb33e905c93396db3b57a4ce0efc984cfd2647dcbe9

Observation f6eb4e40-6224-4237-b50d-b3d76dde49b8 · outbound

This paper cites Learning transferable visual models from natural language supervision.

Stepping Out of Similar Semantic Space for Open-Vocabulary Segmentation Learning transferable visual models from natural language supervision

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:50:24.248952Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T23:50:23.470029Z digest=sha256:e0f13fa3b01274060b894a71f9a0aefc2dcc9858ef61f2bcd6cd0e2533ef5720

Observation ccf2a863-c91f-49d6-a381-782b87cdf801 · outbound

This paper cites Hierarchical Memory for Long Video QA.

Stepping Out of Similar Semantic Space for Open-Vocabulary Segmentation Hierarchical Memory for Long Video QA

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-06T23:50:23.474555Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:50:23.474555Z digest=sha256:51b780be1a6b28e6fed82a85d1ba6cafe94082d2ae974d452af08421a8e23c03

Observation 191b1b35-0da6-4e33-8a0d-247b3a060ea4 · outbound

This paper cites Uni-adafocus: Spatial- temporal dynamic computation for video recognition.

Stepping Out of Similar Semantic Space for Open-Vocabulary Segmentation Uni-adafocus: Spatial- temporal dynamic computation for video recognition

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:50:24.233510Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T23:50:23.479341Z digest=sha256:66fe6dba8b0406c88a31a0a2adb421d231470fba2ec3a523edba81499606bbc2

Observation 8be1fb6b-ddb6-4466-92f6-3e7cc8b144e0 · outbound

This paper cites Iterprime: Zero-shot referring image segmen- tation with iterative grad-cam refinement and primary word emphasis.

Stepping Out of Similar Semantic Space for Open-Vocabulary Segmentation Iterprime: Zero-shot referring image segmen- tation with iterative grad-cam refinement and primary word emphasis

Reference 49

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:50:24.218075Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T23:50:23.484430Z digest=sha256:2107142ac7d9ab50ddcbd13ff217c6607ddd8d7118434484a9847cbf26b6f885

Observation cb973aa9-23da-4116-afe2-42bef671ea2b · outbound

This paper cites Sam2-love: Segment anything model 2 in language- aided audio-visual scenes.

Stepping Out of Similar Semantic Space for Open-Vocabulary Segmentation Sam2-love: Segment anything model 2 in language- aided audio-visual scenes

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:50:24.202554Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T23:50:23.489538Z digest=sha256:ff5e116a2478ebc7dfb71505b92d3e3218b9905abb73d0e84ae7b7a9940f817a

Observation 494fb9ba-fec2-4b55-8b91-96c449cbf13d · outbound

This paper cites HyperSeg: Towards Universal Visual Segmentation with Large Language Model.

Stepping Out of Similar Semantic Space for Open-Vocabulary Segmentation HyperSeg: Towards Universal Visual Segmentation with Large Language Model

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-06T23:50:23.494720Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:50:23.494720Z digest=sha256:e5341549fba4388ee1325fc85efa4c68c8af4c8f889a2b142d6b184b6cfd7739

Observation 0cb8fa4d-6035-476f-ae78-1a5d6d762ba8 · outbound

This paper cites InstructSeg: Unifying Instructed Visual Segmentation with Multi-modal Large Language Models.

Stepping Out of Similar Semantic Space for Open-Vocabulary Segmentation InstructSeg: Unifying Instructed Visual Segmentation with Multi-modal Large Language Models

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-06T23:50:23.499385Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:50:23.499385Z digest=sha256:feeb3599dcb11dfa5e7271f7be0921b38ac3de70e8983c61210a36a0106ee02b

Observation 6c4d208a-bab4-40b6-90d7-116debeb9aed · outbound

This paper cites A large-scale benchmark for food im- age segmentation.

Stepping Out of Similar Semantic Space for Open-Vocabulary Segmentation A large-scale benchmark for food im- age segmentation

Reference 53

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:50:24.187846Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T23:50:23.504013Z digest=sha256:cef050e766dbe3d7052c74a34ed942241f7950c221eb9426522345613dd0e274

Observation e17d219c-1cdf-4ab9-a8b5-4882523f970d · outbound

This paper cites Detectron2.

Stepping Out of Similar Semantic Space for Open-Vocabulary Segmentation Detectron2

Reference 54

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:50:24.170493Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T23:50:23.508334Z digest=sha256:3a4d9fa158047881805c208841e508603063a67258460b3c4d1b12c6e2e39daf

Observation 6f439054-5bbb-428e-959b-d363152575d9 · outbound

This paper cites Semantic projection network for zero- and few-label semantic segmentation.

Stepping Out of Similar Semantic Space for Open-Vocabulary Segmentation Semantic projection network for zero- and few-label semantic segmentation

Reference 55

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:50:24.155465Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T23:50:23.512584Z digest=sha256:782421f94e3b5a5274f21d9aaf70470e70aae9f0dd06e532e04023931185be9a

Observation 8541a39c-81cf-4695-a0d3-dce7b50db631 · outbound

This paper cites Bridging the gap: A unified video comprehension framework for moment retrieval and highlight detection.

Stepping Out of Similar Semantic Space for Open-Vocabulary Segmentation Bridging the gap: A unified video comprehension framework for moment retrieval and highlight detection

Reference 56

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:50:24.140251Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T23:50:23.516744Z digest=sha256:3e31530375b9ff15cc080e93884ebf9b10719765576604eb23f3b77e8b6e9e4a

Observation a8f54fe3-8ea2-4cab-997b-ab0e6d579725 · outbound

This paper cites Sed: A simple encoder-decoder for open- vocabulary semantic segmentation.

Stepping Out of Similar Semantic Space for Open-Vocabulary Segmentation Sed: A simple encoder-decoder for open- vocabulary semantic segmentation

Reference 57

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:50:24.124526Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T23:50:23.521827Z digest=sha256:6aa0e1e1349f03433dd50b65a48c5a8decfd81af37e90fa9b6b0788ab764a7e8

Observation fe3547b4-b4dd-4f37-83c1-97cfd2ad475c · outbound

This paper cites Alvarez, and Ping Luo.

Stepping Out of Similar Semantic Space for Open-Vocabulary Segmentation Alvarez, and Ping Luo

Reference 58

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:50:24.109620Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T23:50:23.526934Z digest=sha256:1be25b8d764f3a33592e80022fe21589928c9f6afcece5598bec061de0571686

Observation fe505afc-f573-406b-8518-9bbc306c6719 · outbound

This paper cites Open-vocabulary panop- tic segmentation with text-to-image diffusion models.

Stepping Out of Similar Semantic Space for Open-Vocabulary Segmentation Open-vocabulary panop- tic segmentation with text-to-image diffusion models

Reference 59

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:50:24.094278Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T23:50:23.531413Z digest=sha256:c9f97dd082034e44836cf68b42d53be0287dff5c7cd483ad19fc6b8b65e839b3

Observation 17813763-b4c7-4cc1-a2ea-c59571aa80ca · outbound

This paper cites A Simple Baseline for Open-Vocabulary Semantic Segmentation with Pre-trained Vision-language Model.

Stepping Out of Similar Semantic Space for Open-Vocabulary Segmentation A Simple Baseline for Open-Vocabulary Semantic Segmentation with Pre-trained Vision-language Model

Reference 60

Resolution
verified exact
local_arxiv, observed 2026-08-06T23:50:23.626640Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T23:50:23.536291Z digest=sha256:5412336738263604b200e38180d91cfe52cdb2dcca1b7dbce15610a645991f37

Observation cd68d2cc-bf5d-4dde-bcd9-7e618daee1d8 · outbound

This paper cites Side adapter network for open-vocabulary semantic segmentation.

Stepping Out of Similar Semantic Space for Open-Vocabulary Segmentation Side adapter network for open-vocabulary semantic segmentation

Reference 61

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:50:24.077918Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T23:50:23.541877Z digest=sha256:5af1a36b42595b38824686cfd179e971cc6cadf54c0ce2f3dfe3b2f6923e66e6

Observation d0e369b9-683d-46a4-8293-042169a04220 · outbound

This paper cites Masq- clip for open-vocabulary universal image segmentation.

Stepping Out of Similar Semantic Space for Open-Vocabulary Segmentation Masq- clip for open-vocabulary universal image segmentation

Reference 62

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:50:24.061442Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T23:50:23.546341Z digest=sha256:1957897ddf8aa7c67791e8b52f2681bbc50da15add401b2af740253f340736a4

Observation bd2b2bcc-c07e-4916-84ae-123c09abd7b5 · outbound

This paper cites Convolutions die hard: Open-vocabulary seg- mentation with single frozen convolutional clip.

Stepping Out of Similar Semantic Space for Open-Vocabulary Segmentation Convolutions die hard: Open-vocabulary seg- mentation with single frozen convolutional clip

Reference 63

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:50:24.046307Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T23:50:23.550790Z digest=sha256:6a38c52e6c2cb7e40df2bcfe69548237084b03af8f4f1c4087d197e58f3f4413

Observation 0e8d8dd4-b8d1-4016-b6c3-8bb89b40456c · outbound

This paper cites Prototypical matching and open set rejection for zero-shot semantic segmentation.

Stepping Out of Similar Semantic Space for Open-Vocabulary Segmentation Prototypical matching and open set rejection for zero-shot semantic segmentation

Reference 64

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:50:24.024982Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T23:50:23.555457Z digest=sha256:74bbd7a6116402834b9ecf76c4e6e65c8c688d20e77618513bc995b2741bfeb8

Observation fe2be44c-1e16-44a6-ad2d-f8e543fb0667 · outbound

This paper cites Flash-VStream: Memory-Based Real-Time Understanding for Long Video Streams.

Stepping Out of Similar Semantic Space for Open-Vocabulary Segmentation Flash-VStream: Memory-Based Real-Time Understanding for Long Video Streams

Reference 65

Resolution
unresolved
no resolver link, observed 2026-08-06T23:50:23.560907Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:50:23.560907Z digest=sha256:990990e2a0bb18ca069e75b1eccb92e192812f37ac4299ac31e16f7ceb5c1bc1

Observation d2d361ab-ea73-4d7a-9d59-648b450a27b6 · outbound

This paper cites Scene parsing through ADE20K dataset.

Stepping Out of Similar Semantic Space for Open-Vocabulary Segmentation Scene parsing through ADE20K dataset

Reference 66

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:50:24.002543Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T23:50:23.565560Z digest=sha256:b4efc1073b411c13704ba7079e22a3da23b8e644e18e54372a92fe40a53f347f

Pith citing papers

No inbound Pith citation observations are available.