Pith. sign in

Paper Citation Record · LEDGER

Towards Open-Vocabulary Video Semantic Segmentation

As of 17 August 2026, this Paper Citation Record lists 60 of 60 outbound references and 0 inbound Pith citation observations for arXiv:2412.09329.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2412.09329 v1

Coverage vector

measured 60 of 60 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-11T17:11:26.749841Z

measured 60 of 60 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-16T06:30:59.297886+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

60 of 60 outbound references displayed

  • verified exact0
  • verified fuzzy53
  • unresolved7
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 81e86c29-4f9d-4375-a079-dae5ecb5d4d8 · outbound

This paper cites Mining contextual information beyond image for semantic segmentation,.

Towards Open-Vocabulary Video Semantic Segmentation Mining contextual information beyond image for semantic segmentation,

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T17:11:27.427742Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-11T17:11:26.523458Z digest=sha256:27447b25bf1b9ebfed91efbea325f34f185ab088ae07d58e8661ef673ff96b43

Observation 02f48fa2-06de-4412-b765-e1057c8f8ada · outbound

This paper cites ISNet: Integrate image-level and semantic-level context for semantic seg- mentation,.

Towards Open-Vocabulary Video Semantic Segmentation ISNet: Integrate image-level and semantic-level context for semantic seg- mentation,

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T17:11:27.414580Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-11T17:11:26.527818Z digest=sha256:9d8daf306b15f9e8ccbd56f749ef8a53af2c2ffa9c10373865e0d7631d80760d

Observation 7a96571f-faa7-4855-859d-a541140e2086 · outbound

This paper cites Boundary-guided lightweight semantic segmen- tation with multi-scale semantic context,.

Towards Open-Vocabulary Video Semantic Segmentation Boundary-guided lightweight semantic segmen- tation with multi-scale semantic context,

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T17:11:27.402676Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-11T17:11:26.531421Z digest=sha256:000d746abbdd237c6b81f6c817e3b7507161b8ad703a52e25e0aa3641767661c

Observation c3033095-cafe-43cb-8909-59150ceef491 · outbound

This paper cites Fbsnet: A fast bilateral symmetrical network for real-time se- mantic segmentation,.

Towards Open-Vocabulary Video Semantic Segmentation Fbsnet: A fast bilateral symmetrical network for real-time se- mantic segmentation,

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T17:11:27.391161Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-11T17:11:26.535174Z digest=sha256:b0209ef55647d878eb888c3d7beeb78b380a1c64808facbe75b4259540b6074b

Observation ad32c5fe-0071-4c61-bfb8-ae549fd86b3a · outbound

This paper cites Semantic segmen- tation guided pixel fusion for image retargeting,.

Towards Open-Vocabulary Video Semantic Segmentation Semantic segmen- tation guided pixel fusion for image retargeting,

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T17:11:27.378207Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-11T17:11:26.539049Z digest=sha256:baf36912ddded0a57a762732cfc6b292706ea32b6e3a5cbe48fee357f23b4ebd

Observation 9489911a-7756-4592-bdcf-4b2c6b00ea9f · outbound

This paper cites Difference-aware distillation for semantic segmenta- tion,.

Towards Open-Vocabulary Video Semantic Segmentation Difference-aware distillation for semantic segmenta- tion,

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T17:11:27.367388Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-11T17:11:26.542589Z digest=sha256:5eff2f3f7e64aca7e0f9a2a3bd02a5210ad13e388d8aec9a015e98fb000d2218

Observation 1a665bc3-41a8-4314-b182-8101cea5ea32 · outbound

This paper cites Learning transferable visual models from natural language supervision,.

Towards Open-Vocabulary Video Semantic Segmentation Learning transferable visual models from natural language supervision,

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T17:11:27.356619Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-11T17:11:26.546378Z digest=sha256:17341f9f71af42d9f293f8d7bf2a73249364afd4cb142c2649fc0b64796ce77c

Observation 92b5a785-89f4-4373-8e9b-9b3cd9cfefd9 · outbound

This paper cites Side adapter network for open-vocabulary semantic segmen- tation,.

Towards Open-Vocabulary Video Semantic Segmentation Side adapter network for open-vocabulary semantic segmen- tation,

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T17:11:27.344907Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-11T17:11:26.549781Z digest=sha256:a7994ea5f4eb67cd53fe551c2ee73a6480630a1fdecae66378a126dc6175a2ae

Observation 52e28d88-2909-48a0-ab9a-0f60cb700556 · outbound

This paper cites Scaling open- vocabulary image segmentation with image-level labels,.

Towards Open-Vocabulary Video Semantic Segmentation Scaling open- vocabulary image segmentation with image-level labels,

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T17:11:27.333113Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-11T17:11:26.553084Z digest=sha256:dac284aa452aefab8d83fca9bf61c039650ec880d5247fc1c3ffb2442349e973

Observation dbfa2416-1278-46c3-98df-369ad52ff9e7 · outbound

This paper cites Towards open-vocabulary video instance segmentation,.

Towards Open-Vocabulary Video Semantic Segmentation Towards open-vocabulary video instance segmentation,

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T17:11:27.321825Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-11T17:11:26.556384Z digest=sha256:9924db4b1bfd95c202648f6aa41d5d50c72ac50a9f75d2a3c2fdefb6763e98e6

Observation 97bd1120-1f65-421a-8216-a681559d0913 · outbound

This paper cites OpenVIS: Open-vocabulary Video Instance Segmentation.

Towards Open-Vocabulary Video Semantic Segmentation OpenVIS: Open-vocabulary Video Instance Segmentation

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-11T17:11:26.560061Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T17:11:26.560061Z digest=sha256:54a5ba481b38e69e0799786a077ad3cf954616ed239e3860563c7ffa6c71bc11

Observation d61852ce-00dd-40ff-a977-e673ca7de199 · outbound

This paper cites VSPW: A large-scale dataset for video scene parsing in the wild,.

Towards Open-Vocabulary Video Semantic Segmentation VSPW: A large-scale dataset for video scene parsing in the wild,

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T17:11:27.309997Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-11T17:11:26.563955Z digest=sha256:287f9a7ded5169da496439522c248358a86489023df6eea252faf4d86a4f7d23

Observation 282073e0-4599-4ea9-8e45-9c2a8ec6390b · outbound

This paper cites The Cityscapes dataset for semantic urban scene under- standing,.

Towards Open-Vocabulary Video Semantic Segmentation The Cityscapes dataset for semantic urban scene under- standing,

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T17:11:27.299331Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-11T17:11:26.567553Z digest=sha256:0cf695bb4a07b7abc8dba3650850f78e0b3662c6ae51727de11903b7f938abac

Observation 4824f420-a459-4b02-acc6-a5a73e959148 · outbound

This paper cites In- door segmentation and support inference from RGBD images,.

Towards Open-Vocabulary Video Semantic Segmentation In- door segmentation and support inference from RGBD images,

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T17:11:27.288361Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-11T17:11:26.570822Z digest=sha256:3eb253e4bb6357662d738c521543943fd902e0d7b993bf027894de9b5721199e

Observation 8d4b902f-e5ac-4b23-806b-ec93d8658613 · outbound

This paper cites Segmentation and recognition using structure from mo- tion point clouds,.

Towards Open-Vocabulary Video Semantic Segmentation Segmentation and recognition using structure from mo- tion point clouds,

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T17:11:27.277157Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-11T17:11:26.574234Z digest=sha256:447636ed4384ab8f48ff5982812a7f22f5caf94011caca1c35e9ee5293d0cfd5

Observation dd886d01-766c-44a3-9af8-060cad5a563c · outbound

This paper cites Low-latency video semantic segmentation,.

Towards Open-Vocabulary Video Semantic Segmentation Low-latency video semantic segmentation,

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T17:11:27.266496Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-11T17:11:26.578371Z digest=sha256:62c30f823dd522c24f844251f362aa914ea8456a14472f7de31a6d47496d7066

Observation 2fafca02-3be7-48f2-b584-ec1f41f3f764 · outbound

This paper cites Clockwork convnets for video semantic segmentation,.

Towards Open-Vocabulary Video Semantic Segmentation Clockwork convnets for video semantic segmentation,

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T17:11:27.255692Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-11T17:11:26.582186Z digest=sha256:7cabd29148272fb6606e5f70b654c99893d402267eda3b79973bc86645b0d565

Observation 8971abdd-0c8b-4a85-8b67-4feaba13395d · outbound

This paper cites Budget-aware deep semantic video segmentation,.

Towards Open-Vocabulary Video Semantic Segmentation Budget-aware deep semantic video segmentation,

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T17:11:27.245194Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-11T17:11:26.585829Z digest=sha256:8df8bf3a8748756c9fea9d613b8f27f0f2dc144fd311d3edbd6049fb3e8b79e0

Observation afa76a1f-43b7-4f35-9d66-3cef598ae2e4 · outbound

This paper cites Deep feature flow for video recognition,.

Towards Open-Vocabulary Video Semantic Segmentation Deep feature flow for video recognition,

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T17:11:27.234774Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-11T17:11:26.589750Z digest=sha256:45a60d6384e0513b8fb84bf5d7a15b419bff1c4f5fa089cef4414f177c6f379d

Observation 6102fd91-2f2c-42d2-980f-19539f5335af · outbound

This paper cites Dynamic video segmentation network,.

Towards Open-Vocabulary Video Semantic Segmentation Dynamic video segmentation network,

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T17:11:27.224028Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-11T17:11:26.594807Z digest=sha256:37068d08b7fd23ba20e08b02655aed3e317a1fb6cfa8cbd0c6f48574cbab4b0f

Observation 0ceae0fd-845a-448b-b78d-4e7fdab0667c · outbound

This paper cites Accel: A correc- tive fusion network for efficient semantic segmentation on video,.

Towards Open-Vocabulary Video Semantic Segmentation Accel: A correc- tive fusion network for efficient semantic segmentation on video,

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T17:11:27.213728Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-11T17:11:26.598832Z digest=sha256:3f36948b786e13a2c3568684d37e1af3222875e29c85fed4a07e088f75b3b86b

Observation e7fb2c42-0d9b-4388-bf31-aa1013de3212 · outbound

This paper cites Temporally distributed networks for fast video semantic segmentation,.

Towards Open-Vocabulary Video Semantic Segmentation Temporally distributed networks for fast video semantic segmentation,

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T17:11:27.203054Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-11T17:11:26.603256Z digest=sha256:d8a040b5b5b1dd4504f006a99e6e608ef8106b23cf7f23d53c5da719ba2de207

Observation d5a48f03-a61f-492e-aa25-2c388d58f4e7 · outbound

This paper cites Efficient se- mantic video segmentation with per-frame inference,.

Towards Open-Vocabulary Video Semantic Segmentation Efficient se- mantic video segmentation with per-frame inference,

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T17:11:27.191657Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-11T17:11:26.607013Z digest=sha256:1fdba4fe531d12d3372cbcd46e0e4bafcbbbcbe63143d24d56ebfb76aac3ed0e

Observation 7a4c6c64-4370-4914-b0fe-5cced1b01a02 · outbound

This paper cites Local memory attention for fast video semantic seg- mentation,.

Towards Open-Vocabulary Video Semantic Segmentation Local memory attention for fast video semantic seg- mentation,

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T17:11:27.181279Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-11T17:11:26.610719Z digest=sha256:5645e2376427775b20103969dad5f2297577e8522de7f71a695f7d9f15f1fa04

Observation 627b6002-6dcb-4eb8-9502-908c6b054a83 · outbound

This paper cites Feature space optimization for semantic video segmentation,.

Towards Open-Vocabulary Video Semantic Segmentation Feature space optimization for semantic video segmentation,

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T17:11:27.170003Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-11T17:11:26.614271Z digest=sha256:5b0e81338b6838fa3ca7c84fc5bc6df2563d090501de61dd7d1548b252a12433

Observation 93237ddc-f16d-4e40-8aa1-221174f1d06c · outbound

This paper cites Video semantic segmentation via sparse tem- poral transformer,.

Towards Open-Vocabulary Video Semantic Segmentation Video semantic segmentation via sparse tem- poral transformer,

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T17:11:27.159078Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-11T17:11:26.618146Z digest=sha256:9fa04ccec4116438c268aacd20e41a51743d0d65769a0bf87963047751067e61

Observation dde8f402-7a74-4a75-aa1e-c65e96f1099a · outbound

This paper cites Semantic video CNNs through representation warping,.

Towards Open-Vocabulary Video Semantic Segmentation Semantic video CNNs through representation warping,

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T17:11:27.149016Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-11T17:11:26.621869Z digest=sha256:19612742cda8cc7dbea9d45ecba47ee0b0521b2ef54e57f8335f83a8d0edbcea

Observation 68fb74be-67ad-4c62-a6c2-f6e687da86dd · outbound

This paper cites Video scene parsing with predictive feature learning,.

Towards Open-Vocabulary Video Semantic Segmentation Video scene parsing with predictive feature learning,

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T17:11:27.137843Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-11T17:11:26.625700Z digest=sha256:f5fb841f0150773fe3093abc1d71093a0a31290b951978c7c7b7f9cc433b51ff

Observation c0bc9b50-84cc-47c8-b08e-3e274fafa66a · outbound

This paper cites Surveillance video parsing with single frame supervi- sion,.

Towards Open-Vocabulary Video Semantic Segmentation Surveillance video parsing with single frame supervi- sion,

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T17:11:27.127238Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-11T17:11:26.629783Z digest=sha256:74a3e717ecdc726487bf045bdc985529adb591a39c3ad58a71435a529288c6b0

Observation a6e1c533-0069-4787-8762-0c08c3cc94ea · outbound

This paper cites Semantic video seg- mentation by gated recurrent flow propagation,.

Towards Open-Vocabulary Video Semantic Segmentation Semantic video seg- mentation by gated recurrent flow propagation,

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T17:11:27.115836Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-11T17:11:26.634349Z digest=sha256:974a7b1e3fc154554c643bfea117fc2cc70ed2c226010f7514efd046a2e3441c

Observation 974c22ee-1672-4d82-80fe-4d68c1df5601 · outbound

This paper cites Improving semantic segmen- tation via video propagation and label relaxation,.

Towards Open-Vocabulary Video Semantic Segmentation Improving semantic segmen- tation via video propagation and label relaxation,

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T17:11:27.105727Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-11T17:11:26.638204Z digest=sha256:178677b5bfaa6b49cb78c3b3d145c2905f99b43f00cba0f5eba91a4f1ec71110

Observation d3c4eea2-d354-4d0b-a792-42882983972e · outbound

This paper cites AuxAdapt: Stable and efficient test-time adaptation for temporally consistent video semantic segmentation,.

Towards Open-Vocabulary Video Semantic Segmentation AuxAdapt: Stable and efficient test-time adaptation for temporally consistent video semantic segmentation,

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T17:11:27.095049Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-11T17:11:26.642027Z digest=sha256:edd5620bdab1a1911a0d40660c3535f3aa3de12731060579735c266fa1534104

Observation bae41cb6-1f9a-4a52-a2cf-a08c6b0bba00 · outbound

This paper cites Coarse-to-fine feature mining for video semantic seg- mentation,.

Towards Open-Vocabulary Video Semantic Segmentation Coarse-to-fine feature mining for video semantic seg- mentation,

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T17:11:27.084007Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-11T17:11:26.645691Z digest=sha256:0c184496a857755b682905cfbd8d6f52258f3aa38cfc51fbe531104498d5ef9e

Observation 9fa7ab79-35c9-46c2-b22a-146ef82a19de · outbound

This paper cites Mining relations among cross-frame affini- ties for video semantic segmentation,.

Towards Open-Vocabulary Video Semantic Segmentation Mining relations among cross-frame affini- ties for video semantic segmentation,

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T17:11:27.073163Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-11T17:11:26.649497Z digest=sha256:bcb6b82b924684aecb479990e0d71d626208a44183450aa20dd1cd044e4e67df

Observation 319d3c95-ce45-448a-a97a-3dcd08a8c798 · outbound

This paper cites Learning local and global temporal contexts for video semantic segmentation,.

Towards Open-Vocabulary Video Semantic Segmentation Learning local and global temporal contexts for video semantic segmentation,

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T17:11:27.061996Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-11T17:11:26.653524Z digest=sha256:0d26cd52537d649640b1075b736f1febca180c5a092486f80cd0b3764df1992a

Observation 26e07b58-8010-41b1-bd01-809b9d85c177 · outbound

This paper cites Video K-Net: A simple, strong, and unified baseline for video segmentation,.

Towards Open-Vocabulary Video Semantic Segmentation Video K-Net: A simple, strong, and unified baseline for video segmentation,

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T17:11:27.049959Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-11T17:11:26.658184Z digest=sha256:66dad8f847eef5c39fccf34cbebe36f4e1f910a00075a94213f1e12a8c5aa6ad

Observation 86cac5d9-1dd0-430d-ada2-cd7461dafbec · outbound

This paper cites Tube-Link: A flexible cross tube baseline for universal video segmentation,.

Towards Open-Vocabulary Video Semantic Segmentation Tube-Link: A flexible cross tube baseline for universal video segmentation,

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T17:11:27.038709Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-11T17:11:26.661899Z digest=sha256:b9ee2898e1b2bdd9e7de644deabc5d265cd08d8df28c19f475a6603b7e1cfd08

Observation 5ebbae45-4990-4a75-9983-7a924569947e · outbound

This paper cites Mask propagation for efficient video seman- tic segmentation,.

Towards Open-Vocabulary Video Semantic Segmentation Mask propagation for efficient video seman- tic segmentation,

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T17:11:27.026447Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-11T17:11:26.666155Z digest=sha256:f3d27c46f77d883b57f4d7b422ef1547f9c5f9a65176f0546ec7b8e32eeb24d9

Observation 66cedfee-721d-4778-ac7a-aec0adae826e · outbound

This paper cites Scaling up visual and vision-language representation learning with noisy text supervision,.

Towards Open-Vocabulary Video Semantic Segmentation Scaling up visual and vision-language representation learning with noisy text supervision,

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T17:11:27.013685Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-11T17:11:26.670412Z digest=sha256:67b9625be3f7bc5fa50bb19fb17a1342b9188b401d5dc026b5edac8aef6dc447

Observation 1bdc92d7-7239-468d-8113-5974d44d47e1 · outbound

This paper cites Decoupling zero-shot semantic segmentation,.

Towards Open-Vocabulary Video Semantic Segmentation Decoupling zero-shot semantic segmentation,

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-11T17:11:26.674080Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T17:11:26.674080Z digest=sha256:31e9f000157c5c7818ea8732be059efcd48eee2e2d8e0e868efc0b09e9934e0b

Observation 4c9221c0-6dc8-428d-8271-a52fcf48f837 · outbound

This paper cites P2T: Pyramid pooling transformer for scene understanding,.

Towards Open-Vocabulary Video Semantic Segmentation P2T: Pyramid pooling transformer for scene understanding,

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T17:11:26.995835Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-11T17:11:26.677839Z digest=sha256:e548114fe56ca5c4a5f33509399881656313cad4ffbde100b9e76b30f1124041

Observation db07c70f-059a-4257-85fd-e52b1514ada0 · outbound

This paper cites Object-contextual rep- resentations for semantic segmentation,.

Towards Open-Vocabulary Video Semantic Segmentation Object-contextual rep- resentations for semantic segmentation,

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T17:11:26.985143Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-11T17:11:26.681230Z digest=sha256:2d38e76492ab0c9fc7fa0d998d97699497079a1061d399d2f8deb366b1222488

Observation 28507564-4460-4729-9720-ad8a0df8f6ff · outbound

This paper cites Attention is all you need,.

Towards Open-Vocabulary Video Semantic Segmentation Attention is all you need,

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T17:11:26.973924Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-11T17:11:26.684503Z digest=sha256:9ae4147a7e0d2dceab638395953a66f4cec4d25c45fff73825abc04aa1307509

Observation aa28fe8f-a11d-441a-abdd-de22ef3fe6bc · outbound

This paper cites CAT-Seg: Cost Aggregation for Open-Vocabulary Semantic Segmentation.

Towards Open-Vocabulary Video Semantic Segmentation CAT-Seg: Cost Aggregation for Open-Vocabulary Semantic Segmentation

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-11T17:11:26.688063Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T17:11:26.688063Z digest=sha256:e415cf2c478bd6f7138bfbeb161e44b55ff17790dbf0bcff2f177204145d221c

Observation 0028ff7d-c385-4069-85b8-3d9b94ff78eb · outbound

This paper cites Open- vocabulary semantic segmentation with decoupled one- pass network,.

Towards Open-Vocabulary Video Semantic Segmentation Open- vocabulary semantic segmentation with decoupled one- pass network,

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-11T17:11:26.691746Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T17:11:26.691746Z digest=sha256:699eed940f3cf8cd7f4ec3bd176ee882109d6f3896b410c4a2cb186f6a0cb470

Observation 0557ca17-5af9-486d-bbe7-079b0499f56c · outbound

This paper cites A simple baseline for open-vocabulary semantic segmentation with pre-trained vision-language model,.

Towards Open-Vocabulary Video Semantic Segmentation A simple baseline for open-vocabulary semantic segmentation with pre-trained vision-language model,

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T17:11:26.955341Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-11T17:11:26.695206Z digest=sha256:7ef054062d23012971049d7094a9c7a5041cd1f3077942bff8e654537b0039d5

Observation f977d459-4636-4c00-bba0-451543cccecd · outbound

This paper cites FreeSeg: Unified, universal and open-vocabulary image segmentation,.

Towards Open-Vocabulary Video Semantic Segmentation FreeSeg: Unified, universal and open-vocabulary image segmentation,

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T17:11:26.944054Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-11T17:11:26.698823Z digest=sha256:d095ef08e31b76f0f3c774303baf5ef7b68e74b9a3d39e4e4aa1e51e558e155d

Observation 9aa7eb27-0aee-4dd2-a39b-ff521d0c6ac8 · outbound

This paper cites Segment anything,.

Towards Open-Vocabulary Video Semantic Segmentation Segment anything,

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T17:11:26.932089Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-11T17:11:26.702201Z digest=sha256:7bfbc040537c3cc86d43d33425f720e648defde297db09ecdbe1673a5502797d

Observation 16387f5a-0b5d-4845-ac12-08d25bc9cf07 · outbound

This paper cites An image is worth 16x16 words: Transformers for image recognition at scale,.

Towards Open-Vocabulary Video Semantic Segmentation An image is worth 16x16 words: Transformers for image recognition at scale,

Reference 49

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T17:11:26.920788Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-11T17:11:26.706197Z digest=sha256:d98b84efbb636fa71420911e9dfc8617e97ed5c6f3b6f7cd345f6a549d193b60

Observation 8c8fbd71-1638-4165-8c6f-6bc1c68a14f2 · outbound

This paper cites Deep residual learning for image recognition,.

Towards Open-Vocabulary Video Semantic Segmentation Deep residual learning for image recognition,

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-11T17:11:26.710110Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T17:11:26.710110Z digest=sha256:fbe127b8b39a5c0f581388266bfb23b6720d95b2fa838f79bc57ecaeca4407bf

Observation 0497bd56-09bb-41fb-ab76-7a9f35eade14 · outbound

This paper cites Adam: A method for stochas- tic optimization,.

Towards Open-Vocabulary Video Semantic Segmentation Adam: A method for stochas- tic optimization,

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T17:11:26.903013Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-11T17:11:26.714856Z digest=sha256:4f5cb24811848400f80d1f7cbdb872da2f52244baddce92bdd53e4e436eb93ee

Observation 5fdec607-f165-4b1d-a85e-0b71df270525 · outbound

This paper cites Deep-irtarget: An automatic target detector in infrared imagery using dual-domain feature extraction and allo- cation,.

Towards Open-Vocabulary Video Semantic Segmentation Deep-irtarget: An automatic target detector in infrared imagery using dual-domain feature extraction and allo- cation,

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T17:11:26.891752Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-11T17:11:26.719034Z digest=sha256:c43e1aaa1f658d2171d3132ad4380c23af8a74e6473a34f2086002442cf55ce4

Observation a7ca01f1-4d8c-4ed9-95dd-910ed3c26d0c · outbound

This paper cites Semantic scene completion from a single depth image,.

Towards Open-Vocabulary Video Semantic Segmentation Semantic scene completion from a single depth image,

Reference 53

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T17:11:26.878108Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-11T17:11:26.722768Z digest=sha256:a943d90111ac681a16057896790119c3ad10e7f20f8d27b303247f148c0b7ff6

Observation 1de4735a-cdf3-4795-b0b2-d142ba9224e3 · outbound

This paper cites Con- trastive boundary learning for point cloud segmentation,.

Towards Open-Vocabulary Video Semantic Segmentation Con- trastive boundary learning for point cloud segmentation,

Reference 54

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T17:11:26.867046Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-11T17:11:26.726425Z digest=sha256:1c6f8b0eafad5e53baa06d55e3a9e724db8ba09fa6aac58b52649689a1625a9c

Observation a1f9b628-a1a2-46c7-b622-df85e84be200 · outbound

This paper cites Learning from noisy labels with deep neural networks: A survey,.

Towards Open-Vocabulary Video Semantic Segmentation Learning from noisy labels with deep neural networks: A survey,

Reference 55

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T17:11:26.855396Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-11T17:11:26.730497Z digest=sha256:ccd8bb8bab267b4a201cc240d4909e54dd3f5772f9cd2f3180a75f05f8fb2016

Observation 972b680e-8ebe-4456-a322-ca9b6909f5e3 · outbound

This paper cites Cognition-driven structural prior for instance- dependent label transition matrix estimation,.

Towards Open-Vocabulary Video Semantic Segmentation Cognition-driven structural prior for instance- dependent label transition matrix estimation,

Reference 56

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T17:11:26.843706Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-11T17:11:26.733962Z digest=sha256:b3f082c40db4812b866b1496fb25b1fd122f84bf1713ac89b5ba34b0647bda03

Observation 3b2d0864-d636-4836-bcc0-ac34a8d7c856 · outbound

This paper cites Feature modulation transformer: Cross-refinement of global representation via high-frequency prior for image super-resolution,.

Towards Open-Vocabulary Video Semantic Segmentation Feature modulation transformer: Cross-refinement of global representation via high-frequency prior for image super-resolution,

Reference 57

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T17:11:26.832274Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-11T17:11:26.737738Z digest=sha256:ad0d6e67f04dec022c3094ab8a3b026634e8ca687f4823140bece7fe548ccaa6

Observation 4e297a28-a321-41b2-969d-9fee4b2c15d5 · outbound

This paper cites Panet: Few-shot image semantic segmentation with prototype alignment,.

Towards Open-Vocabulary Video Semantic Segmentation Panet: Few-shot image semantic segmentation with prototype alignment,

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-11T17:11:26.741789Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T17:11:26.741789Z digest=sha256:c174feee19b012c44c4161a27b4316efb4229cc9223d6709c7ff66621619cf36

Observation dd3fc7a2-3f14-4046-8a4c-f174a650ab0b · outbound

This paper cites Part-aware correlation networks for few-shot learning,.

Towards Open-Vocabulary Video Semantic Segmentation Part-aware correlation networks for few-shot learning,

Reference 59

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T17:11:26.814259Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-11T17:11:26.745666Z digest=sha256:9754b68eb7e59c5a3d1ca963285fde328e072992106a6093e4546e7b83294df9

Observation 77c0a1a2-cea4-43d1-902e-030e8f973aab · outbound

This paper cites an unresolved cited work.

Towards Open-Vocabulary Video Semantic Segmentation Unresolved cited work

Reference 1998

Resolution
unresolved
raw_fallback, observed 2026-08-11T17:11:26.802409Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-11T17:11:26.749841Z digest=sha256:50714d77679eb05aeae7c629b2982871c1299b81d90442fb6a18b41727e1f4cb

Pith citing papers

No inbound Pith citation observations are available.