Pith. sign in

Paper Citation Record · LEDGER

Towards Open-Vocabulary Video Semantic Segmentation

As of 17 August 2026, this Paper Citation Record lists 60 of 60 outbound references and 0 inbound Pith citation observations for arXiv:2412.09329.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2412.09329 v1

Coverage vector

measured 60 of 60 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-11T17:11:26.749841Z

measured 60 of 60 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-17T06:30:58.91139+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

60 of 60 outbound references displayed

  • verified exact0
  • verified fuzzy53
  • unresolved7
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 81e86c29-4f9d-4375-a079-dae5ecb5d4d8 · outbound

This paper cites Mining contextual information beyond image for semantic segmentation,.

Towards Open-Vocabulary Video Semantic Segmentation Mining contextual information beyond image for semantic segmentation,

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T17:11:27.427742Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-11T17:11:26.523458Z digest=sha256:a90f767946495dbb9b3439e285a336364039c24226163c4b8c5a0bef2bfd18db

Observation 02f48fa2-06de-4412-b765-e1057c8f8ada · outbound

This paper cites ISNet: Integrate image-level and semantic-level context for semantic seg- mentation,.

Towards Open-Vocabulary Video Semantic Segmentation ISNet: Integrate image-level and semantic-level context for semantic seg- mentation,

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T17:11:27.414580Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-11T17:11:26.527818Z digest=sha256:3436506afb804387c3f97cb1f53e791415e5de7f65b4bdf36a7ce7b4be78697c

Observation 7a96571f-faa7-4855-859d-a541140e2086 · outbound

This paper cites Boundary-guided lightweight semantic segmen- tation with multi-scale semantic context,.

Towards Open-Vocabulary Video Semantic Segmentation Boundary-guided lightweight semantic segmen- tation with multi-scale semantic context,

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T17:11:27.402676Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-11T17:11:26.531421Z digest=sha256:cf93f64fe4039abe4368f60228851ae48fdec8fd37eeb7f0a2bff5c052903fd4

Observation c3033095-cafe-43cb-8909-59150ceef491 · outbound

This paper cites Fbsnet: A fast bilateral symmetrical network for real-time se- mantic segmentation,.

Towards Open-Vocabulary Video Semantic Segmentation Fbsnet: A fast bilateral symmetrical network for real-time se- mantic segmentation,

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T17:11:27.391161Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-11T17:11:26.535174Z digest=sha256:218dbe70f2e9b29a3e0b19ffac38f0ecc1fd357ccfa7c5062710a00c966ce1f8

Observation ad32c5fe-0071-4c61-bfb8-ae549fd86b3a · outbound

This paper cites Semantic segmen- tation guided pixel fusion for image retargeting,.

Towards Open-Vocabulary Video Semantic Segmentation Semantic segmen- tation guided pixel fusion for image retargeting,

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T17:11:27.378207Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-11T17:11:26.539049Z digest=sha256:274beee4633084887bf3f3fed117bd6913fbb0b0092be6fb1735067ac044b260

Observation 9489911a-7756-4592-bdcf-4b2c6b00ea9f · outbound

This paper cites Difference-aware distillation for semantic segmenta- tion,.

Towards Open-Vocabulary Video Semantic Segmentation Difference-aware distillation for semantic segmenta- tion,

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T17:11:27.367388Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-11T17:11:26.542589Z digest=sha256:2759676c6cb88ecf7e8651cbdfbc83c1639c988076f826739c9b1e88f3a72465

Observation 1a665bc3-41a8-4314-b182-8101cea5ea32 · outbound

This paper cites Learning transferable visual models from natural language supervision,.

Towards Open-Vocabulary Video Semantic Segmentation Learning transferable visual models from natural language supervision,

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T17:11:27.356619Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-11T17:11:26.546378Z digest=sha256:edc195c920467fb40c6e7231ab5c5a95f7c58cd68191ab5fa8c332f3d4778048

Observation 92b5a785-89f4-4373-8e9b-9b3cd9cfefd9 · outbound

This paper cites Side adapter network for open-vocabulary semantic segmen- tation,.

Towards Open-Vocabulary Video Semantic Segmentation Side adapter network for open-vocabulary semantic segmen- tation,

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T17:11:27.344907Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-11T17:11:26.549781Z digest=sha256:f3a696a72ec07c38daecaa90333bcca87f1afe78f8a4713024aa1a3aa2225f76

Observation 52e28d88-2909-48a0-ab9a-0f60cb700556 · outbound

This paper cites Scaling open- vocabulary image segmentation with image-level labels,.

Towards Open-Vocabulary Video Semantic Segmentation Scaling open- vocabulary image segmentation with image-level labels,

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T17:11:27.333113Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-11T17:11:26.553084Z digest=sha256:e583b58d86ce6e2a4f2e7d15e9b3862e4c40bdec53afd0b667076f7de7384f76

Observation dbfa2416-1278-46c3-98df-369ad52ff9e7 · outbound

This paper cites Towards open-vocabulary video instance segmentation,.

Towards Open-Vocabulary Video Semantic Segmentation Towards open-vocabulary video instance segmentation,

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T17:11:27.321825Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-11T17:11:26.556384Z digest=sha256:bde151a0def91e93c0b7b731e778d1af4cfcde25b42ba9ab10fdfcc4b8258ee6

Observation 97bd1120-1f65-421a-8216-a681559d0913 · outbound

This paper cites OpenVIS: Open-vocabulary Video Instance Segmentation.

Towards Open-Vocabulary Video Semantic Segmentation OpenVIS: Open-vocabulary Video Instance Segmentation

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-11T17:11:26.560061Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T17:11:26.560061Z digest=sha256:54a5ba481b38e69e0799786a077ad3cf954616ed239e3860563c7ffa6c71bc11

Observation d61852ce-00dd-40ff-a977-e673ca7de199 · outbound

This paper cites VSPW: A large-scale dataset for video scene parsing in the wild,.

Towards Open-Vocabulary Video Semantic Segmentation VSPW: A large-scale dataset for video scene parsing in the wild,

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T17:11:27.309997Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-11T17:11:26.563955Z digest=sha256:67427339dc5787e81e8124a189784c0116700984c28c6a87c00479a5e84eb3cf

Observation 282073e0-4599-4ea9-8e45-9c2a8ec6390b · outbound

This paper cites The Cityscapes dataset for semantic urban scene under- standing,.

Towards Open-Vocabulary Video Semantic Segmentation The Cityscapes dataset for semantic urban scene under- standing,

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T17:11:27.299331Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-11T17:11:26.567553Z digest=sha256:05e8fd53595ad3db28696f6e36cf0b1f2aeb54ce3473a874c6e033ba403ae059

Observation 4824f420-a459-4b02-acc6-a5a73e959148 · outbound

This paper cites In- door segmentation and support inference from RGBD images,.

Towards Open-Vocabulary Video Semantic Segmentation In- door segmentation and support inference from RGBD images,

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T17:11:27.288361Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-11T17:11:26.570822Z digest=sha256:96d9cf92f53df9818472551f4ae7262dea7c725f4885914932b6e0ccc4a084e5

Observation 8d4b902f-e5ac-4b23-806b-ec93d8658613 · outbound

This paper cites Segmentation and recognition using structure from mo- tion point clouds,.

Towards Open-Vocabulary Video Semantic Segmentation Segmentation and recognition using structure from mo- tion point clouds,

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T17:11:27.277157Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-11T17:11:26.574234Z digest=sha256:3803cdd357220de0ae6464ae52149aeffc68fcfb0bb90743303581c1a4a943dd

Observation dd886d01-766c-44a3-9af8-060cad5a563c · outbound

This paper cites Low-latency video semantic segmentation,.

Towards Open-Vocabulary Video Semantic Segmentation Low-latency video semantic segmentation,

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T17:11:27.266496Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-11T17:11:26.578371Z digest=sha256:af63d5f6a686e66df7c30bb85c02654481a624c6b551f7875197758a66552c09

Observation 2fafca02-3be7-48f2-b584-ec1f41f3f764 · outbound

This paper cites Clockwork convnets for video semantic segmentation,.

Towards Open-Vocabulary Video Semantic Segmentation Clockwork convnets for video semantic segmentation,

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T17:11:27.255692Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-11T17:11:26.582186Z digest=sha256:38d7919552437d3003910b897c7063232903cf7d059760dbffffbc7c81f440f4

Observation 8971abdd-0c8b-4a85-8b67-4feaba13395d · outbound

This paper cites Budget-aware deep semantic video segmentation,.

Towards Open-Vocabulary Video Semantic Segmentation Budget-aware deep semantic video segmentation,

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T17:11:27.245194Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-11T17:11:26.585829Z digest=sha256:19fb7e70ddb94e4bbd928549c8be7105c839c5bb1ab14459ae9435750d377c39

Observation afa76a1f-43b7-4f35-9d66-3cef598ae2e4 · outbound

This paper cites Deep feature flow for video recognition,.

Towards Open-Vocabulary Video Semantic Segmentation Deep feature flow for video recognition,

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T17:11:27.234774Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-11T17:11:26.589750Z digest=sha256:719b1537d4fca6d3d1d83d15eb16ecb3cb5183ed3e9ae929b79aebadf5b84bd4

Observation 6102fd91-2f2c-42d2-980f-19539f5335af · outbound

This paper cites Dynamic video segmentation network,.

Towards Open-Vocabulary Video Semantic Segmentation Dynamic video segmentation network,

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T17:11:27.224028Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-11T17:11:26.594807Z digest=sha256:385ad48b704522e75ce342ebc3b62fb5c1413bedb035db785548535174fe4c09

Observation 0ceae0fd-845a-448b-b78d-4e7fdab0667c · outbound

This paper cites Accel: A correc- tive fusion network for efficient semantic segmentation on video,.

Towards Open-Vocabulary Video Semantic Segmentation Accel: A correc- tive fusion network for efficient semantic segmentation on video,

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T17:11:27.213728Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-11T17:11:26.598832Z digest=sha256:8e31e061ed7ac77d0a6458ac464e626c382ebcd78477471557d07511b5e130a3

Observation e7fb2c42-0d9b-4388-bf31-aa1013de3212 · outbound

This paper cites Temporally distributed networks for fast video semantic segmentation,.

Towards Open-Vocabulary Video Semantic Segmentation Temporally distributed networks for fast video semantic segmentation,

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T17:11:27.203054Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-11T17:11:26.603256Z digest=sha256:b9c655877d5cba135d499086bc3caaddf16830ef726691007f946255fadb3919

Observation d5a48f03-a61f-492e-aa25-2c388d58f4e7 · outbound

This paper cites Efficient se- mantic video segmentation with per-frame inference,.

Towards Open-Vocabulary Video Semantic Segmentation Efficient se- mantic video segmentation with per-frame inference,

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T17:11:27.191657Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-11T17:11:26.607013Z digest=sha256:586f52b9c642bc0a8a2ffe14ab4a9729440980fd8a007aaf06df22de01143bb9

Observation 7a4c6c64-4370-4914-b0fe-5cced1b01a02 · outbound

This paper cites Local memory attention for fast video semantic seg- mentation,.

Towards Open-Vocabulary Video Semantic Segmentation Local memory attention for fast video semantic seg- mentation,

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T17:11:27.181279Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-11T17:11:26.610719Z digest=sha256:19f59389946b4a103b98a8da8818494b1ac8ac876f32278bd3bce6476aca7114

Observation 627b6002-6dcb-4eb8-9502-908c6b054a83 · outbound

This paper cites Feature space optimization for semantic video segmentation,.

Towards Open-Vocabulary Video Semantic Segmentation Feature space optimization for semantic video segmentation,

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T17:11:27.170003Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-11T17:11:26.614271Z digest=sha256:8ef143654e134008731a56e73c6d2c8c241c211f0f62e604965af8c0b1ae8fc4

Observation 93237ddc-f16d-4e40-8aa1-221174f1d06c · outbound

This paper cites Video semantic segmentation via sparse tem- poral transformer,.

Towards Open-Vocabulary Video Semantic Segmentation Video semantic segmentation via sparse tem- poral transformer,

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T17:11:27.159078Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-11T17:11:26.618146Z digest=sha256:fc96eb9957bd44678e62d5d35b2fc3126343a053abe21d8068ebd922820069d3

Observation dde8f402-7a74-4a75-aa1e-c65e96f1099a · outbound

This paper cites Semantic video CNNs through representation warping,.

Towards Open-Vocabulary Video Semantic Segmentation Semantic video CNNs through representation warping,

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T17:11:27.149016Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-11T17:11:26.621869Z digest=sha256:19bc22096cc02d4edb885a0b9d539438bf1661eb867b2e0550e2f7d3f199db1d

Observation 68fb74be-67ad-4c62-a6c2-f6e687da86dd · outbound

This paper cites Video scene parsing with predictive feature learning,.

Towards Open-Vocabulary Video Semantic Segmentation Video scene parsing with predictive feature learning,

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T17:11:27.137843Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-11T17:11:26.625700Z digest=sha256:e5d2b3ffe6b801d4445c0803ee6baba443c88f99c498611ac3b99e247cf7f0ee

Observation c0bc9b50-84cc-47c8-b08e-3e274fafa66a · outbound

This paper cites Surveillance video parsing with single frame supervi- sion,.

Towards Open-Vocabulary Video Semantic Segmentation Surveillance video parsing with single frame supervi- sion,

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T17:11:27.127238Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-11T17:11:26.629783Z digest=sha256:ef87f3673449119c00cd49292dae0914dc5919e8e81e5fba8c5eb59e7c004cc9

Observation a6e1c533-0069-4787-8762-0c08c3cc94ea · outbound

This paper cites Semantic video seg- mentation by gated recurrent flow propagation,.

Towards Open-Vocabulary Video Semantic Segmentation Semantic video seg- mentation by gated recurrent flow propagation,

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T17:11:27.115836Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-11T17:11:26.634349Z digest=sha256:106de762fb56354e82b755bd698311ff9e8650b3343adba2972b62d886bfc1b1

Observation 974c22ee-1672-4d82-80fe-4d68c1df5601 · outbound

This paper cites Improving semantic segmen- tation via video propagation and label relaxation,.

Towards Open-Vocabulary Video Semantic Segmentation Improving semantic segmen- tation via video propagation and label relaxation,

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T17:11:27.105727Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-11T17:11:26.638204Z digest=sha256:48f7eebdba0a646c201b55dec4db76678c6856b6d0e7c4ecf39cd6ea70d15ac7

Observation d3c4eea2-d354-4d0b-a792-42882983972e · outbound

This paper cites AuxAdapt: Stable and efficient test-time adaptation for temporally consistent video semantic segmentation,.

Towards Open-Vocabulary Video Semantic Segmentation AuxAdapt: Stable and efficient test-time adaptation for temporally consistent video semantic segmentation,

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T17:11:27.095049Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-11T17:11:26.642027Z digest=sha256:0784e488c5bbddfbef2893fac5e8cf5454645b360ed3f6d5781c5fd8ffbafea6

Observation bae41cb6-1f9a-4a52-a2cf-a08c6b0bba00 · outbound

This paper cites Coarse-to-fine feature mining for video semantic seg- mentation,.

Towards Open-Vocabulary Video Semantic Segmentation Coarse-to-fine feature mining for video semantic seg- mentation,

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T17:11:27.084007Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-11T17:11:26.645691Z digest=sha256:f89928a079696f716414c75173d75ae28ac0a441df3b5d0d011a5d84f1b0fe36

Observation 9fa7ab79-35c9-46c2-b22a-146ef82a19de · outbound

This paper cites Mining relations among cross-frame affini- ties for video semantic segmentation,.

Towards Open-Vocabulary Video Semantic Segmentation Mining relations among cross-frame affini- ties for video semantic segmentation,

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T17:11:27.073163Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-11T17:11:26.649497Z digest=sha256:8cb282d9306c4ae93fa04fec8eeff9bb178f62aae294f2e49c31bf9d3f6b73db

Observation 319d3c95-ce45-448a-a97a-3dcd08a8c798 · outbound

This paper cites Learning local and global temporal contexts for video semantic segmentation,.

Towards Open-Vocabulary Video Semantic Segmentation Learning local and global temporal contexts for video semantic segmentation,

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T17:11:27.061996Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-11T17:11:26.653524Z digest=sha256:dc7f0094dc6365e2cbbce29d7402b39cabbd515a7815a0efe7cec2c843578658

Observation 26e07b58-8010-41b1-bd01-809b9d85c177 · outbound

This paper cites Video K-Net: A simple, strong, and unified baseline for video segmentation,.

Towards Open-Vocabulary Video Semantic Segmentation Video K-Net: A simple, strong, and unified baseline for video segmentation,

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T17:11:27.049959Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-11T17:11:26.658184Z digest=sha256:a9347605fbe602414e80b7b650559d1c826167ad57ae7d8eb6125bbe5eeaf3e6

Observation 86cac5d9-1dd0-430d-ada2-cd7461dafbec · outbound

This paper cites Tube-Link: A flexible cross tube baseline for universal video segmentation,.

Towards Open-Vocabulary Video Semantic Segmentation Tube-Link: A flexible cross tube baseline for universal video segmentation,

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T17:11:27.038709Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-11T17:11:26.661899Z digest=sha256:1a1db0c34b2f7a2422b8d4c02e55f932aae405765350f31b25b7819401bc9a7b

Observation 5ebbae45-4990-4a75-9983-7a924569947e · outbound

This paper cites Mask propagation for efficient video seman- tic segmentation,.

Towards Open-Vocabulary Video Semantic Segmentation Mask propagation for efficient video seman- tic segmentation,

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T17:11:27.026447Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-11T17:11:26.666155Z digest=sha256:6d1092d01014d8c1ae92696f16848cadcb4a81c4bba9ed694f8867d3ba7676fc

Observation 66cedfee-721d-4778-ac7a-aec0adae826e · outbound

This paper cites Scaling up visual and vision-language representation learning with noisy text supervision,.

Towards Open-Vocabulary Video Semantic Segmentation Scaling up visual and vision-language representation learning with noisy text supervision,

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T17:11:27.013685Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-11T17:11:26.670412Z digest=sha256:0c4be984b7aa69f38ed5e11d6cd51739ea37bb6f028273fa028083a19930acd5

Observation 1bdc92d7-7239-468d-8113-5974d44d47e1 · outbound

This paper cites Decoupling zero-shot semantic segmentation,.

Towards Open-Vocabulary Video Semantic Segmentation Decoupling zero-shot semantic segmentation,

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-11T17:11:26.674080Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T17:11:26.674080Z digest=sha256:31e9f000157c5c7818ea8732be059efcd48eee2e2d8e0e868efc0b09e9934e0b

Observation 4c9221c0-6dc8-428d-8271-a52fcf48f837 · outbound

This paper cites P2T: Pyramid pooling transformer for scene understanding,.

Towards Open-Vocabulary Video Semantic Segmentation P2T: Pyramid pooling transformer for scene understanding,

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T17:11:26.995835Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-11T17:11:26.677839Z digest=sha256:b25f0976d8360612112641e73b74b3c09c80fa70015bf523d426292bf03513bd

Observation db07c70f-059a-4257-85fd-e52b1514ada0 · outbound

This paper cites Object-contextual rep- resentations for semantic segmentation,.

Towards Open-Vocabulary Video Semantic Segmentation Object-contextual rep- resentations for semantic segmentation,

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T17:11:26.985143Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-11T17:11:26.681230Z digest=sha256:5d67d0fe8c82fc57d46b11fe0bf4e4e760ac958bbf71f1e45d4aab92ae1bc304

Observation 28507564-4460-4729-9720-ad8a0df8f6ff · outbound

This paper cites Attention is all you need,.

Towards Open-Vocabulary Video Semantic Segmentation Attention is all you need,

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T17:11:26.973924Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-11T17:11:26.684503Z digest=sha256:e7a0aedf704bf0341cdd181e97d6c2b20f70d886529fbe194a2124c3556207f0

Observation aa28fe8f-a11d-441a-abdd-de22ef3fe6bc · outbound

This paper cites CAT-Seg: Cost Aggregation for Open-Vocabulary Semantic Segmentation.

Towards Open-Vocabulary Video Semantic Segmentation CAT-Seg: Cost Aggregation for Open-Vocabulary Semantic Segmentation

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-11T17:11:26.688063Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T17:11:26.688063Z digest=sha256:e415cf2c478bd6f7138bfbeb161e44b55ff17790dbf0bcff2f177204145d221c

Observation 0028ff7d-c385-4069-85b8-3d9b94ff78eb · outbound

This paper cites Open- vocabulary semantic segmentation with decoupled one- pass network,.

Towards Open-Vocabulary Video Semantic Segmentation Open- vocabulary semantic segmentation with decoupled one- pass network,

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-11T17:11:26.691746Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T17:11:26.691746Z digest=sha256:699eed940f3cf8cd7f4ec3bd176ee882109d6f3896b410c4a2cb186f6a0cb470

Observation 0557ca17-5af9-486d-bbe7-079b0499f56c · outbound

This paper cites A simple baseline for open-vocabulary semantic segmentation with pre-trained vision-language model,.

Towards Open-Vocabulary Video Semantic Segmentation A simple baseline for open-vocabulary semantic segmentation with pre-trained vision-language model,

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T17:11:26.955341Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-11T17:11:26.695206Z digest=sha256:2ca4f7baf090e367b299035317c9baf3d1f4a7d4cb72491eb61f36c7cf7f5c95

Observation f977d459-4636-4c00-bba0-451543cccecd · outbound

This paper cites FreeSeg: Unified, universal and open-vocabulary image segmentation,.

Towards Open-Vocabulary Video Semantic Segmentation FreeSeg: Unified, universal and open-vocabulary image segmentation,

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T17:11:26.944054Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-11T17:11:26.698823Z digest=sha256:687b34c9553fef3caba7b52c347e791d9a60a8b0ab11f6d1ddf515d3e08669b1

Observation 9aa7eb27-0aee-4dd2-a39b-ff521d0c6ac8 · outbound

This paper cites Segment anything,.

Towards Open-Vocabulary Video Semantic Segmentation Segment anything,

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T17:11:26.932089Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-11T17:11:26.702201Z digest=sha256:c122a73efcd71c860d02c0f00c936ca43ea68daae12acefe76324ef3778a5fbf

Observation 16387f5a-0b5d-4845-ac12-08d25bc9cf07 · outbound

This paper cites An image is worth 16x16 words: Transformers for image recognition at scale,.

Towards Open-Vocabulary Video Semantic Segmentation An image is worth 16x16 words: Transformers for image recognition at scale,

Reference 49

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T17:11:26.920788Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-11T17:11:26.706197Z digest=sha256:cde541e7ce9cb0267703e5e5cdc1ebf22a8bc3ba9b73dcd1d0139948f2725ed5

Observation 8c8fbd71-1638-4165-8c6f-6bc1c68a14f2 · outbound

This paper cites Deep residual learning for image recognition,.

Towards Open-Vocabulary Video Semantic Segmentation Deep residual learning for image recognition,

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-11T17:11:26.710110Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T17:11:26.710110Z digest=sha256:fbe127b8b39a5c0f581388266bfb23b6720d95b2fa838f79bc57ecaeca4407bf

Observation 0497bd56-09bb-41fb-ab76-7a9f35eade14 · outbound

This paper cites Adam: A method for stochas- tic optimization,.

Towards Open-Vocabulary Video Semantic Segmentation Adam: A method for stochas- tic optimization,

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T17:11:26.903013Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-11T17:11:26.714856Z digest=sha256:5d96eb04077ebb23575f579ab2542af4d55a3736b5d40c0137ba8d8215353b94

Observation 5fdec607-f165-4b1d-a85e-0b71df270525 · outbound

This paper cites Deep-irtarget: An automatic target detector in infrared imagery using dual-domain feature extraction and allo- cation,.

Towards Open-Vocabulary Video Semantic Segmentation Deep-irtarget: An automatic target detector in infrared imagery using dual-domain feature extraction and allo- cation,

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T17:11:26.891752Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-11T17:11:26.719034Z digest=sha256:8ffffc8d0316345046344969be2c6553d0e5fc2606fa70f6c80ec5357da7dfe3

Observation a7ca01f1-4d8c-4ed9-95dd-910ed3c26d0c · outbound

This paper cites Semantic scene completion from a single depth image,.

Towards Open-Vocabulary Video Semantic Segmentation Semantic scene completion from a single depth image,

Reference 53

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T17:11:26.878108Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-11T17:11:26.722768Z digest=sha256:2ba4932cc3fb3a4da6b5d4822d1b2a9e59bdb14d10a0d57128f3b17717abc5e7

Observation 1de4735a-cdf3-4795-b0b2-d142ba9224e3 · outbound

This paper cites Con- trastive boundary learning for point cloud segmentation,.

Towards Open-Vocabulary Video Semantic Segmentation Con- trastive boundary learning for point cloud segmentation,

Reference 54

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T17:11:26.867046Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-11T17:11:26.726425Z digest=sha256:63ab7633051e308c5eec030952a25e34ad992e821558b87bfd38f99edc1affe4

Observation a1f9b628-a1a2-46c7-b622-df85e84be200 · outbound

This paper cites Learning from noisy labels with deep neural networks: A survey,.

Towards Open-Vocabulary Video Semantic Segmentation Learning from noisy labels with deep neural networks: A survey,

Reference 55

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T17:11:26.855396Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-11T17:11:26.730497Z digest=sha256:42405b24d42c5f3bfe47a936c91b2986e86a802d93a591d5dc4becef304f6384

Observation 972b680e-8ebe-4456-a322-ca9b6909f5e3 · outbound

This paper cites Cognition-driven structural prior for instance- dependent label transition matrix estimation,.

Towards Open-Vocabulary Video Semantic Segmentation Cognition-driven structural prior for instance- dependent label transition matrix estimation,

Reference 56

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T17:11:26.843706Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-11T17:11:26.733962Z digest=sha256:f5000436a65295ffb90c647b8fd093803fe0fa9628f839e10e48b6103fc47c8d

Observation 3b2d0864-d636-4836-bcc0-ac34a8d7c856 · outbound

This paper cites Feature modulation transformer: Cross-refinement of global representation via high-frequency prior for image super-resolution,.

Towards Open-Vocabulary Video Semantic Segmentation Feature modulation transformer: Cross-refinement of global representation via high-frequency prior for image super-resolution,

Reference 57

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T17:11:26.832274Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-11T17:11:26.737738Z digest=sha256:bb72ac61bc5138b5d4bdb8855724fdc5db11f7dd44dc046414e1a7b598860269

Observation 4e297a28-a321-41b2-969d-9fee4b2c15d5 · outbound

This paper cites Panet: Few-shot image semantic segmentation with prototype alignment,.

Towards Open-Vocabulary Video Semantic Segmentation Panet: Few-shot image semantic segmentation with prototype alignment,

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-11T17:11:26.741789Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T17:11:26.741789Z digest=sha256:c174feee19b012c44c4161a27b4316efb4229cc9223d6709c7ff66621619cf36

Observation dd3fc7a2-3f14-4046-8a4c-f174a650ab0b · outbound

This paper cites Part-aware correlation networks for few-shot learning,.

Towards Open-Vocabulary Video Semantic Segmentation Part-aware correlation networks for few-shot learning,

Reference 59

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T17:11:26.814259Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-11T17:11:26.745666Z digest=sha256:c2faf533ddde51891235ac1b770fa60ef049f0869a9d82c562065c5365e0636e

Observation 77c0a1a2-cea4-43d1-902e-030e8f973aab · outbound

This paper cites an unresolved cited work.

Towards Open-Vocabulary Video Semantic Segmentation Unresolved cited work

Reference 1998

Resolution
unresolved
raw_fallback, observed 2026-08-11T17:11:26.802409Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-11T17:11:26.749841Z digest=sha256:b67515d767a1aba203affb1590b0dd9b93b67cbca717533f7c9c05ee222e675e

Pith citing papers

No inbound Pith citation observations are available.