Pith. sign in

Paper Citation Record · LEDGER

STAC: Selective Spatiotemporal Aggregation and Compression for Video Reasoning Segmentation

As of 7 August 2026, this Paper Citation Record lists 86 of 86 outbound references and 0 inbound Pith citation observations for arXiv:2607.02922.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2607.02922 v1

Coverage vector

measured 86 of 86 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-07-12T06:06:47.233814Z

measured 86 of 86 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

86 of 86 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved86
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation d4f6b78e-f922-41b3-abab-2b27e0411ec9 · outbound

This paper cites GPT-4 Technical Report.

STAC: Selective Spatiotemporal Aggregation and Compression for Video Reasoning Segmentation GPT-4 Technical Report

Reference 1

Resolution
unresolved
no resolver link, observed 2026-07-12T06:06:47.233814Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T06:06:47.233814Z digest=sha256:cab55ab83effe59f11b0843d96c035e1d63fa5ee6c95cd25f447fb5f66f09b3c

Observation fd14e178-fc05-4119-afee-5954dfb61084 · outbound

This paper cites arXiv preprint arXiv:1412.69801412(6) (2014).

STAC: Selective Spatiotemporal Aggregation and Compression for Video Reasoning Segmentation arXiv preprint arXiv:1412.69801412(6) (2014)

Reference 2

Resolution
unresolved
no resolver link, observed 2026-07-12T06:06:47.233814Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T06:06:47.233814Z digest=sha256:cd1acc7a5af46078fb136ea25c9e1d2240f6d170cf9e3cbff70fb2d020aeb6f5

Observation a731ca61-a436-431d-9f5e-44dc3a6c75bb · outbound

This paper cites Machine Intelligence Research (2026).

STAC: Selective Spatiotemporal Aggregation and Compression for Video Reasoning Segmentation Machine Intelligence Research (2026)

Reference 3

Resolution
unresolved
no resolver link, observed 2026-07-12T06:06:47.233814Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T06:06:47.233814Z digest=sha256:27673c99f6672d24282ee04f7b6f875837d47142cb7a30ae29cae31c3df5e8f3

Observation b43e701c-f674-426e-b1be-b193255ae447 · outbound

This paper cites Machine Intelligence Research (2026).

STAC: Selective Spatiotemporal Aggregation and Compression for Video Reasoning Segmentation Machine Intelligence Research (2026)

Reference 4

Resolution
unresolved
no resolver link, observed 2026-07-12T06:06:47.233814Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T06:06:47.233814Z digest=sha256:3830526eec3b25d64a5d4769abfdd8d8c15e4b9a52670e3e9fd9faf7c40def6e

Observation a985ebc9-90f4-4304-a321-f9f016748f5a · outbound

This paper cites Qwen-VL: A Versatile Vision-Language Model for Understanding, Localization, Text Reading, and Beyond.

STAC: Selective Spatiotemporal Aggregation and Compression for Video Reasoning Segmentation Qwen-VL: A Versatile Vision-Language Model for Understanding, Localization, Text Reading, and Beyond

Reference 5

Resolution
unresolved
no resolver link, observed 2026-07-12T06:06:47.233814Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T06:06:47.233814Z digest=sha256:6777fdc301742d3b13b78b187a97a33d06eb3bdf9d3a6e2c70c957f0be08523f

Observation e9344be0-e3fc-440a-bf3a-297b14c07ac1 · outbound

This paper cites In: NeurIPS.

STAC: Selective Spatiotemporal Aggregation and Compression for Video Reasoning Segmentation In: NeurIPS

Reference 6

Resolution
unresolved
no resolver link, observed 2026-07-12T06:06:47.233814Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T06:06:47.233814Z digest=sha256:9388f4b71320aa5212c8eb1dcfabe1ce7fd5e6881d9173d954c9cea9e5f8b981

Observation 43e53513-0d5a-4527-b976-2a3b4b884d8c · outbound

This paper cites Estimating or Propagating Gradients Through Stochastic Neurons for Conditional Computation.

STAC: Selective Spatiotemporal Aggregation and Compression for Video Reasoning Segmentation Estimating or Propagating Gradients Through Stochastic Neurons for Conditional Computation

Reference 7

Resolution
unresolved
no resolver link, observed 2026-07-12T06:06:47.233814Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T06:06:47.233814Z digest=sha256:8e093f153e43ab2bbce0b588935cc8c0edb27ef0182933ac1fa740854411e2c9

Observation 286409b6-3575-4587-9819-987bf083cbab · outbound

This paper cites Token Merging: Your ViT But Faster.

STAC: Selective Spatiotemporal Aggregation and Compression for Video Reasoning Segmentation Token Merging: Your ViT But Faster

Reference 8

Resolution
unresolved
no resolver link, observed 2026-07-12T06:06:47.233814Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T06:06:47.233814Z digest=sha256:d3bf29884e6798ac0e57869f106eb0641db65ef137d3adad32b0765702c9cfac

Observation 49536b96-5fe3-4f53-82ea-8fa6a26f86de · outbound

This paper cites In: CVPR.

STAC: Selective Spatiotemporal Aggregation and Compression for Video Reasoning Segmentation In: CVPR

Reference 9

Resolution
unresolved
no resolver link, observed 2026-07-12T06:06:47.233814Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T06:06:47.233814Z digest=sha256:e193fcc314f3351a5d5c65771c0e331586220be79e65ddb0e9914ceb4a3df14b

Observation 60d65879-347f-4703-ae19-0bdd6bbc6dad · outbound

This paper cites Video Mamba Suite: State Space Model as a Versatile Alternative for Video Understanding.

STAC: Selective Spatiotemporal Aggregation and Compression for Video Reasoning Segmentation Video Mamba Suite: State Space Model as a Versatile Alternative for Video Understanding

Reference 10

Resolution
unresolved
no resolver link, observed 2026-07-12T06:06:47.233814Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T06:06:47.233814Z digest=sha256:065d2b30dbdf23e42ce4d6c1b3cf237cfa304d0e47a89bea9c8f2b865689823a

Observation 2273dcbc-1b9f-4414-873f-eea92e22b690 · outbound

This paper cites In: ECCV.

STAC: Selective Spatiotemporal Aggregation and Compression for Video Reasoning Segmentation In: ECCV

Reference 11

Resolution
unresolved
no resolver link, observed 2026-07-12T06:06:47.233814Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T06:06:47.233814Z digest=sha256:81a150413b697ee5ea09b1afd584cfe59d764a410b7afd68e3ca0fd5d64524c4

Observation e5a4a78c-c2e6-4cd1-b2a9-ad8f673a58f8 · outbound

This paper cites LongVILA: Scaling Long-Context Visual Language Models for Long Videos.

STAC: Selective Spatiotemporal Aggregation and Compression for Video Reasoning Segmentation LongVILA: Scaling Long-Context Visual Language Models for Long Videos

Reference 12

Resolution
unresolved
no resolver link, observed 2026-07-12T06:06:47.233814Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T06:06:47.233814Z digest=sha256:7d30d5c50db71a7694044b1a7c9deedefde9433c3fd1a2fba8f5ae26fe75fb8f

Observation ee914331-38a4-4e45-bce5-deeb1d576497 · outbound

This paper cites InternVL: Scaling up Vision Foundation Models and Aligning for Generic Visual-Linguistic Tasks.

STAC: Selective Spatiotemporal Aggregation and Compression for Video Reasoning Segmentation InternVL: Scaling up Vision Foundation Models and Aligning for Generic Visual-Linguistic Tasks

Reference 13

Resolution
unresolved
no resolver link, observed 2026-07-12T06:06:47.233814Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T06:06:47.233814Z digest=sha256:82024db53b895e2cd935ca6d287574310dfb16b1c072c8410d3aa794e0941786

Observation 2428d069-2089-40df-9b83-285c6c5ecf9e · outbound

This paper cites VideoLLaMA 2: Advancing Spatial-Temporal Modeling and Audio Understanding in Video-LLMs.

STAC: Selective Spatiotemporal Aggregation and Compression for Video Reasoning Segmentation VideoLLaMA 2: Advancing Spatial-Temporal Modeling and Audio Understanding in Video-LLMs

Reference 14

Resolution
unresolved
no resolver link, observed 2026-07-12T06:06:47.233814Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T06:06:47.233814Z digest=sha256:c5eb2ad7cf0fc7850988982d3a032750505e01214be94e5cfd9bf080b1e72d06

Observation 7c353ec4-af22-4fbb-a93c-8a486d8cfdcc · outbound

This paper cites Transformers are SSMs: Generalized Models and Efficient Algorithms Through Structured State Space Duality.

STAC: Selective Spatiotemporal Aggregation and Compression for Video Reasoning Segmentation Transformers are SSMs: Generalized Models and Efficient Algorithms Through Structured State Space Duality

Reference 15

Resolution
unresolved
no resolver link, observed 2026-07-12T06:06:47.233814Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T06:06:47.233814Z digest=sha256:0686314e8b68ffa68e2a642582ef4877d1668c5061a6666869c4ab9c370d33a2

Observation d83652dd-aaa0-4204-acaf-6e0c470f5855 · outbound

This paper cites In: ICCV.

STAC: Selective Spatiotemporal Aggregation and Compression for Video Reasoning Segmentation In: ICCV

Reference 16

Resolution
unresolved
no resolver link, observed 2026-07-12T06:06:47.233814Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T06:06:47.233814Z digest=sha256:7e486983b637afddbcdd25408b6606f9d33b826889379a0d6c4e3a68ff4ff575

Observation 86b087dd-01bb-4cd4-a2dd-e66afb5d88c4 · outbound

This paper cites In: CVPR.

STAC: Selective Spatiotemporal Aggregation and Compression for Video Reasoning Segmentation In: CVPR

Reference 17

Resolution
unresolved
no resolver link, observed 2026-07-12T06:06:47.233814Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T06:06:47.233814Z digest=sha256:831bf692d6a0ca4ea451f3405607695250ace6ed464b3f44128c06e694856ab6

Observation f6cd77ae-5788-4da2-ae8e-044d04b1e68c · outbound

This paper cites arXiv e-prints pp.

STAC: Selective Spatiotemporal Aggregation and Compression for Video Reasoning Segmentation arXiv e-prints pp

Reference 18

Resolution
unresolved
no resolver link, observed 2026-07-12T06:06:47.233814Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T06:06:47.233814Z digest=sha256:bede8b7c03aa5de42e3fa81a67ed1cd42ad255b63f496aeb01a96f5cf17efb04

Observation 58980e28-8142-4b70-ada0-6d47a6e9d725 · outbound

This paper cites Machine Intelligence Research (2026).

STAC: Selective Spatiotemporal Aggregation and Compression for Video Reasoning Segmentation Machine Intelligence Research (2026)

Reference 19

Resolution
unresolved
no resolver link, observed 2026-07-12T06:06:47.233814Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T06:06:47.233814Z digest=sha256:df865cc625c03e8f6d31caef4748e97d894199e5926d2e6a7bf7c6fa02bb5d5a

Observation 839d4958-a5a6-42d6-90d3-5e302b800bc3 · outbound

This paper cites In: CVPR.

STAC: Selective Spatiotemporal Aggregation and Compression for Video Reasoning Segmentation In: CVPR

Reference 20

Resolution
unresolved
no resolver link, observed 2026-07-12T06:06:47.233814Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T06:06:47.233814Z digest=sha256:9736a373e92b6961007ce521d23d8f917894be2efdbf53f40831b9bbef1acc09

Observation 28d79346-3a6e-4693-ba8c-432a47844da1 · outbound

This paper cites In: Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition.

STAC: Selective Spatiotemporal Aggregation and Compression for Video Reasoning Segmentation In: Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition

Reference 21

Resolution
unresolved
no resolver link, observed 2026-07-12T06:06:47.233814Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T06:06:47.233814Z digest=sha256:d727acf66618ba7126c3d9ef60aa3dec436718005f2c78d5a1c6252515effb7c

Observation 6d6050d8-1e45-4a09-b003-b4fc68496778 · outbound

This paper cites In: First Conference on Language Modeling (2024).

STAC: Selective Spatiotemporal Aggregation and Compression for Video Reasoning Segmentation In: First Conference on Language Modeling (2024)

Reference 22

Resolution
unresolved
no resolver link, observed 2026-07-12T06:06:47.233814Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T06:06:47.233814Z digest=sha256:792ad73344689ca8390b374664a9cda2b53124cf7ab3000745a2e3c291244571

Observation cf3a40f1-d2d0-42db-a3b5-6523e2ea6d9c · outbound

This paper cites In: NeurIPS.

STAC: Selective Spatiotemporal Aggregation and Compression for Video Reasoning Segmentation In: NeurIPS

Reference 23

Resolution
unresolved
no resolver link, observed 2026-07-12T06:06:47.233814Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T06:06:47.233814Z digest=sha256:43715070e4539e2a6a916bc1aed6f320ba4f4b7ca053c7beea4d69dfc7650fb2

Observation 13095465-ae21-4a06-a09f-14e954dfa0d5 · outbound

This paper cites Efficiently Modeling Long Sequences with Structured State Spaces.

STAC: Selective Spatiotemporal Aggregation and Compression for Video Reasoning Segmentation Efficiently Modeling Long Sequences with Structured State Spaces

Reference 24

Resolution
unresolved
no resolver link, observed 2026-07-12T06:06:47.233814Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T06:06:47.233814Z digest=sha256:30a808947e4335dfd1c9a30e82cadcc28b66afb8f2b1ec4d982dd868d6fd4f07

Observation 39dc3970-31d6-4ab3-8225-8dac56440536 · outbound

This paper cites In: ICLR (2022).

STAC: Selective Spatiotemporal Aggregation and Compression for Video Reasoning Segmentation In: ICLR (2022)

Reference 25

Resolution
unresolved
no resolver link, observed 2026-07-12T06:06:47.233814Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T06:06:47.233814Z digest=sha256:7c017c4979f3aac56fc91041d1666e53d31443cf0ba4c2b10c16243c2e5d5e47

Observation 8be4e71c-fc84-401f-abc4-486ced89d171 · outbound

This paper cites In: NeurIPS.

STAC: Selective Spatiotemporal Aggregation and Compression for Video Reasoning Segmentation In: NeurIPS

Reference 26

Resolution
unresolved
no resolver link, observed 2026-07-12T06:06:47.233814Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T06:06:47.233814Z digest=sha256:7bd741ec03b902f4d44a5ccca603a397757a3575ace5d6bf5e92aa7f9b488ca2

Observation 7f46380b-9f22-48d3-85d4-b3e7202a9ec6 · outbound

This paper cites In: NeurIPS.

STAC: Selective Spatiotemporal Aggregation and Compression for Video Reasoning Segmentation In: NeurIPS

Reference 27

Resolution
unresolved
no resolver link, observed 2026-07-12T06:06:47.233814Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T06:06:47.233814Z digest=sha256:e9c705117a9bef0dc26be254ca74f63f75f479e8a976220dae215d456cee3c99

Observation 91864828-bab9-45e9-9bf7-a21edaed84d4 · outbound

This paper cites Machine Intelligence Research (2026) STAC 17.

STAC: Selective Spatiotemporal Aggregation and Compression for Video Reasoning Segmentation Machine Intelligence Research (2026) STAC 17

Reference 28

Resolution
unresolved
no resolver link, observed 2026-07-12T06:06:47.233814Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T06:06:47.233814Z digest=sha256:a6fc7434459eb0c0d6b9aba66504a8819401feda652b1c738259388f98af599a

Observation 0b37b6cf-bf96-4d8b-b8d6-9dacca63d7c1 · outbound

This paper cites In: CVPR.

STAC: Selective Spatiotemporal Aggregation and Compression for Video Reasoning Segmentation In: CVPR

Reference 29

Resolution
unresolved
no resolver link, observed 2026-07-12T06:06:47.233814Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T06:06:47.233814Z digest=sha256:5df073fa64bc9523cdf42d10a418b05abbe086015783759ab2771e2bdc8fa03a

Observation 5f7d59a9-f7df-4e15-b2c1-8a903d976df7 · outbound

This paper cites In: ICLR (2022).

STAC: Selective Spatiotemporal Aggregation and Compression for Video Reasoning Segmentation In: ICLR (2022)

Reference 30

Resolution
unresolved
no resolver link, observed 2026-07-12T06:06:47.233814Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T06:06:47.233814Z digest=sha256:b3ddf6a8e604376b1f9d3a8e8cb7f1258a6ca0b62da1f3ab61287455faf65ad7

Observation 66d601ba-6c07-44d3-af1e-3ff8c8562a84 · outbound

This paper cites In: CVPR.

STAC: Selective Spatiotemporal Aggregation and Compression for Video Reasoning Segmentation In: CVPR

Reference 31

Resolution
unresolved
no resolver link, observed 2026-07-12T06:06:47.233814Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T06:06:47.233814Z digest=sha256:fac64c776ab69b921444023e6d2019cd4ffbe49c831c3f7d90cfb7169811f2fd

Observation 5434a24e-294f-4761-a24c-9514f4e5b6b0 · outbound

This paper cites In: CVPR.

STAC: Selective Spatiotemporal Aggregation and Compression for Video Reasoning Segmentation In: CVPR

Reference 32

Resolution
unresolved
no resolver link, observed 2026-07-12T06:06:47.233814Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T06:06:47.233814Z digest=sha256:311914830a5b07aa7bf7faccc44545e6db53b2bda956296202eb0c9bb8e66a1a

Observation f8435b2d-3802-4302-800f-b4156d23f27b · outbound

This paper cites an unresolved cited work.

STAC: Selective Spatiotemporal Aggregation and Compression for Video Reasoning Segmentation Unresolved cited work

Reference 33

Resolution
unresolved
no resolver link, observed 2026-07-12T06:06:47.233814Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T06:06:47.233814Z digest=sha256:1c7e2ea89d663d7edebb5348887794e9c0a5ce36f5b8728f26fd944a3e2d9dc2

Observation d95f45dd-7816-4adb-838a-ccf224d39e04 · outbound

This paper cites Mixtral of Experts.

STAC: Selective Spatiotemporal Aggregation and Compression for Video Reasoning Segmentation Mixtral of Experts

Reference 34

Resolution
unresolved
no resolver link, observed 2026-07-12T06:06:47.233814Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T06:06:47.233814Z digest=sha256:d172c047eff21f28ec29c9f2c36543145f24f6d7b3868115dd6346d01f290edf

Observation cb8fd7a6-1ec7-4bec-a92a-cc179f284e2d · outbound

This paper cites arXiv preprint arXiv:2503.04130 (2025).

STAC: Selective Spatiotemporal Aggregation and Compression for Video Reasoning Segmentation arXiv preprint arXiv:2503.04130 (2025)

Reference 35

Resolution
unresolved
no resolver link, observed 2026-07-12T06:06:47.233814Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T06:06:47.233814Z digest=sha256:9f684e11980a7747596668fde95a0db07922adad09f0b6cc0fe0531c200a4bd3

Observation b6c65628-7cc4-42a5-80d4-65265b56df39 · outbound

This paper cites Journal of Basic Engineering82(1), 35–45 (1960).

STAC: Selective Spatiotemporal Aggregation and Compression for Video Reasoning Segmentation Journal of Basic Engineering82(1), 35–45 (1960)

Reference 36

Resolution
unresolved
no resolver link, observed 2026-07-12T06:06:47.233814Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T06:06:47.233814Z digest=sha256:b99606054c21a7ed5ef83930595e8f0b06ad872e38a1b863a15418ae245dc03b

Observation 461e5042-f45e-45f8-8ec3-bbba9c00e7e7 · outbound

This paper cites Computational Visual Media11(3), 655–667 (2025).

STAC: Selective Spatiotemporal Aggregation and Compression for Video Reasoning Segmentation Computational Visual Media11(3), 655–667 (2025)

Reference 37

Resolution
unresolved
no resolver link, observed 2026-07-12T06:06:47.233814Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T06:06:47.233814Z digest=sha256:28f21c3803cd7f1fa14e768f3d54e72b33ec31b240a263086fbe2a39f4cd4db8

Observation 07d234a6-566e-4f88-aab3-4927bb33ded9 · outbound

This paper cites In: Asian conference on computer vision.

STAC: Selective Spatiotemporal Aggregation and Compression for Video Reasoning Segmentation In: Asian conference on computer vision

Reference 38

Resolution
unresolved
no resolver link, observed 2026-07-12T06:06:47.233814Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T06:06:47.233814Z digest=sha256:1780fee229b9f22745edb66b3a3f8d0e87d47018903ae097e7eccffa99b6b587

Observation ad9c4785-e53e-4274-88b2-db8d0ad6a5df · outbound

This paper cites In: ICCV.

STAC: Selective Spatiotemporal Aggregation and Compression for Video Reasoning Segmentation In: ICCV

Reference 39

Resolution
unresolved
no resolver link, observed 2026-07-12T06:06:47.233814Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T06:06:47.233814Z digest=sha256:da5d0bc721979caaf7e3ee4dcac89bb87ad632aff74c356fd17876ca861a737e

Observation 50c01312-37ba-4f83-9556-43bc100d8206 · outbound

This paper cites In: ICCV.

STAC: Selective Spatiotemporal Aggregation and Compression for Video Reasoning Segmentation In: ICCV

Reference 40

Resolution
unresolved
no resolver link, observed 2026-07-12T06:06:47.233814Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T06:06:47.233814Z digest=sha256:60e125e02fbd6bdfde5b499297cf1c910f86bb98832ad045d516d3aea21081cc

Observation 44799c7f-5fff-4b2c-ae49-e3e226033daa · outbound

This paper cites In: CVPR.

STAC: Selective Spatiotemporal Aggregation and Compression for Video Reasoning Segmentation In: CVPR

Reference 41

Resolution
unresolved
no resolver link, observed 2026-07-12T06:06:47.233814Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T06:06:47.233814Z digest=sha256:9087e5cdd8b6f8209d26b42cd27c9cd68b658879e398674add607a89c55e23fa

Observation 1a136f9a-79bc-4341-a9d8-7663fd96a47c · outbound

This paper cites LLaVA-OneVision: Easy Visual Task Transfer.

STAC: Selective Spatiotemporal Aggregation and Compression for Video Reasoning Segmentation LLaVA-OneVision: Easy Visual Task Transfer

Reference 42

Resolution
unresolved
no resolver link, observed 2026-07-12T06:06:47.233814Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T06:06:47.233814Z digest=sha256:a18a21b475fef17862904c91b1de3704fe178a6cb021c301d511941b54bea123

Observation 3be3c91e-b7b6-44b9-98fe-723c075550cb · outbound

This paper cites In: ECCV.

STAC: Selective Spatiotemporal Aggregation and Compression for Video Reasoning Segmentation In: ECCV

Reference 43

Resolution
unresolved
no resolver link, observed 2026-07-12T06:06:47.233814Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T06:06:47.233814Z digest=sha256:217e5fc66a093205341b4f659a5217d3bd706957ae88622f0de04f80dabae184

Observation 52ffdd98-43f9-447a-9a6b-11525769d7f8 · outbound

This paper cites an unresolved cited work.

STAC: Selective Spatiotemporal Aggregation and Compression for Video Reasoning Segmentation Unresolved cited work

Reference 44

Resolution
unresolved
no resolver link, observed 2026-07-12T06:06:47.233814Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T06:06:47.233814Z digest=sha256:9dbcbcd0fa810d46591abf413069b4bf4a5d2a3702aaf95b9698e5bf9f89d590

Observation 0aa1c8bd-29c0-4fc9-9186-2e06fef0cd9d · outbound

This paper cites Video-LLaVA: Learning United Visual Representation by Alignment Before Projection.

STAC: Selective Spatiotemporal Aggregation and Compression for Video Reasoning Segmentation Video-LLaVA: Learning United Visual Representation by Alignment Before Projection

Reference 45

Resolution
unresolved
no resolver link, observed 2026-07-12T06:06:47.233814Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T06:06:47.233814Z digest=sha256:0f317fc4fce6e32e8fc5611d928e76b5bd10cd40baaddc766a7a432455d8925e

Observation fca6b943-2e8b-4f21-8027-042bedadfe68 · outbound

This paper cites In: CVPR.

STAC: Selective Spatiotemporal Aggregation and Compression for Video Reasoning Segmentation In: CVPR

Reference 46

Resolution
unresolved
no resolver link, observed 2026-07-12T06:06:47.233814Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T06:06:47.233814Z digest=sha256:c1ab69cddd75c35793b51758b8a5eb20a3304a096eb331cf90764be1055cb220

Observation 19bdcc5a-64fa-4aae-a66c-1a1086274fe5 · outbound

This paper cites an unresolved cited work.

STAC: Selective Spatiotemporal Aggregation and Compression for Video Reasoning Segmentation Unresolved cited work

Reference 47

Resolution
unresolved
no resolver link, observed 2026-07-12T06:06:47.233814Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T06:06:47.233814Z digest=sha256:2f11d0edc85ad00e9b51fedaa88296620eea899766136e5951148378d60109af

Observation e73b5dfc-90ce-4da9-bcaf-296b92770c4d · outbound

This paper cites In: NeurIPS.

STAC: Selective Spatiotemporal Aggregation and Compression for Video Reasoning Segmentation In: NeurIPS

Reference 48

Resolution
unresolved
no resolver link, observed 2026-07-12T06:06:47.233814Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T06:06:47.233814Z digest=sha256:431e9fe28893e87464009a5da27a49fb816f644ba096d52339d357b84f8c8156

Observation fdd9ddd9-0a28-4f01-9f4d-50e2fbfc5cdf · outbound

This paper cites NeurIPS37, 103031–103063 (2024).

STAC: Selective Spatiotemporal Aggregation and Compression for Video Reasoning Segmentation NeurIPS37, 103031–103063 (2024)

Reference 49

Resolution
unresolved
no resolver link, observed 2026-07-12T06:06:47.233814Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T06:06:47.233814Z digest=sha256:9b7c78df5fc4f649e3988d91420ddd09ec06bd9a69046d960c74bd2bd4830ef2

Observation d14f4efd-bb8a-4503-ac09-5e79097118ee · outbound

This paper cites Machine Intelligence Research21(4), 670–683 (2024).

STAC: Selective Spatiotemporal Aggregation and Compression for Video Reasoning Segmentation Machine Intelligence Research21(4), 670–683 (2024)

Reference 50

Resolution
unresolved
no resolver link, observed 2026-07-12T06:06:47.233814Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T06:06:47.233814Z digest=sha256:e0018ff07fe7701ffd1fba2a641eabe8764bbbcfdc61276cc81ddcd231f3d08e

Observation 7b2ce9c1-ab38-4dd4-ba6f-2c5e01d42977 · outbound

This paper cites arXiv e-prints pp.

STAC: Selective Spatiotemporal Aggregation and Compression for Video Reasoning Segmentation arXiv e-prints pp

Reference 51

Resolution
unresolved
no resolver link, observed 2026-07-12T06:06:47.233814Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T06:06:47.233814Z digest=sha256:1bee0d3d064f0efb5e9737bd8e03cdcede87fa41948565613b7dbbfd6b223e3a

Observation fe4f295d-3fd1-4abb-9653-153300840e30 · outbound

This paper cites In: NeurIPS.

STAC: Selective Spatiotemporal Aggregation and Compression for Video Reasoning Segmentation In: NeurIPS

Reference 52

Resolution
unresolved
no resolver link, observed 2026-07-12T06:06:47.233814Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T06:06:47.233814Z digest=sha256:33b7a0ec2bc22fa05916440f82bcab9dfd704e98c1e1bc8f6716e48a0aec62fb

Observation a216cdec-57c7-42b8-b8f9-1d092d8cb6a7 · outbound

This paper cites In: Proceedings of the IEEE/CVF Interna- tional Conference on Computer Vision.

STAC: Selective Spatiotemporal Aggregation and Compression for Video Reasoning Segmentation In: Proceedings of the IEEE/CVF Interna- tional Conference on Computer Vision

Reference 53

Resolution
unresolved
no resolver link, observed 2026-07-12T06:06:47.233814Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T06:06:47.233814Z digest=sha256:047368cec03882e2b6be2d43a35381ff302443032b9ed544fe8eac26e882dcc0

Observation a02ea75c-7acd-46fc-861e-5f796ca732c7 · outbound

This paper cites In: CVPR.

STAC: Selective Spatiotemporal Aggregation and Compression for Video Reasoning Segmentation In: CVPR

Reference 54

Resolution
unresolved
no resolver link, observed 2026-07-12T06:06:47.233814Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T06:06:47.233814Z digest=sha256:8d19fe97dfa8968d5e6db40a0a83a95e9e0a4a2d039b165471ebee7486de396d

Observation 3ce13bc8-a305-4cf6-b2b4-3c4e22d205cd · outbound

This paper cites Computational Visual Media12(1), 71–84 (2026).

STAC: Selective Spatiotemporal Aggregation and Compression for Video Reasoning Segmentation Computational Visual Media12(1), 71–84 (2026)

Reference 55

Resolution
unresolved
no resolver link, observed 2026-07-12T06:06:47.233814Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T06:06:47.233814Z digest=sha256:b3e4efbfc55a47cf7cdaa60297411f42158b0b988c150377b2507d36d039db20

Observation e57e284a-daf4-49c6-8eb7-e0d902de0f7a · outbound

This paper cites In: NeurIPS.

STAC: Selective Spatiotemporal Aggregation and Compression for Video Reasoning Segmentation In: NeurIPS

Reference 56

Resolution
unresolved
no resolver link, observed 2026-07-12T06:06:47.233814Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T06:06:47.233814Z digest=sha256:b1c055bc98d8f561391495db59630bb4ae6c28172cae251287d51abda5320fbd

Observation ce403898-ed40-42f8-a75f-c14eccd175b4 · outbound

This paper cites In: ECCV.

STAC: Selective Spatiotemporal Aggregation and Compression for Video Reasoning Segmentation In: ECCV

Reference 57

Resolution
unresolved
no resolver link, observed 2026-07-12T06:06:47.233814Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T06:06:47.233814Z digest=sha256:d7b22461619197cd2a3858ba0679be603197693055d3eb644d21edb5f20086a5

Observation 99e6b471-2162-4ab1-a085-5dd6ecea0bf9 · outbound

This paper cites In: AAAI.

STAC: Selective Spatiotemporal Aggregation and Compression for Video Reasoning Segmentation In: AAAI

Reference 58

Resolution
unresolved
no resolver link, observed 2026-07-12T06:06:47.233814Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T06:06:47.233814Z digest=sha256:9980f0e9e4674c335a3ca5574bf60e0ba8d42d7a46ba1ce337d34792a3e4abcd

Observation e5a10d60-7622-49ec-be24-312751ef4129 · outbound

This paper cites The 2017 DAVIS Challenge on Video Object Segmentation.

STAC: Selective Spatiotemporal Aggregation and Compression for Video Reasoning Segmentation The 2017 DAVIS Challenge on Video Object Segmentation

Reference 59

Resolution
unresolved
no resolver link, observed 2026-07-12T06:06:47.233814Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T06:06:47.233814Z digest=sha256:86e74bd507c6ea7b5c1829c848cb44b1d2b9483c9db8fe415aa2c99a5613fdd9

Observation 75917728-ec17-4d9f-a731-f863bc70975d · outbound

This paper cites an unresolved cited work.

STAC: Selective Spatiotemporal Aggregation and Compression for Video Reasoning Segmentation Unresolved cited work

Reference 60

Resolution
unresolved
no resolver link, observed 2026-07-12T06:06:47.233814Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T06:06:47.233814Z digest=sha256:33e38bdb71d5f7d557027f026f6386f0ccfc6999a8c9c680598a9444f2f24591

Observation 6575c041-ef40-4a1d-a2bb-b7e44731f92f · outbound

This paper cites In: NeurIPS.

STAC: Selective Spatiotemporal Aggregation and Compression for Video Reasoning Segmentation In: NeurIPS

Reference 61

Resolution
unresolved
no resolver link, observed 2026-07-12T06:06:47.233814Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T06:06:47.233814Z digest=sha256:7dd3567751fa1de310c46caeca2b62adc557a937c977bf17ba045d5aac8e42f1

Observation 622741c0-e100-44a9-9d93-a76f3c5d7186 · outbound

This paper cites In: Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition.

STAC: Selective Spatiotemporal Aggregation and Compression for Video Reasoning Segmentation In: Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition

Reference 62

Resolution
unresolved
no resolver link, observed 2026-07-12T06:06:47.233814Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T06:06:47.233814Z digest=sha256:bd937e579505b0e6ea208b328dfe75a5d557bdab578fdad0774e4e4611d462bb

Observation 20f518b8-8b32-4cc0-9ed3-94ebeef0aa29 · outbound

This paper cites SAM 2: Segment Anything in Images and Videos.

STAC: Selective Spatiotemporal Aggregation and Compression for Video Reasoning Segmentation SAM 2: Segment Anything in Images and Videos

Reference 63

Resolution
unresolved
no resolver link, observed 2026-07-12T06:06:47.233814Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T06:06:47.233814Z digest=sha256:1f1d58fc2958abcd627a2bd70b63ff207d286455f2759b126896ce0bf73db161

Observation 702d9478-8f9f-4295-8b22-b14605d92368 · outbound

This paper cites In: ECCV.

STAC: Selective Spatiotemporal Aggregation and Compression for Video Reasoning Segmentation In: ECCV

Reference 64

Resolution
unresolved
no resolver link, observed 2026-07-12T06:06:47.233814Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T06:06:47.233814Z digest=sha256:634db6e94cd886ee5958311f8bd6a24347af5cc3bc77886d095ae7511e4ce823

Observation 54e0cb18-43af-407f-b79f-dc31dac84aca · outbound

This paper cites TempMe: Video Temporal Token Merging for Efficient Text-Video Retrieval.

STAC: Selective Spatiotemporal Aggregation and Compression for Video Reasoning Segmentation TempMe: Video Temporal Token Merging for Efficient Text-Video Retrieval

Reference 65

Resolution
unresolved
no resolver link, observed 2026-07-12T06:06:47.233814Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T06:06:47.233814Z digest=sha256:fe7ad09f7795b5f45ccb69ef177b9a426d2d12f4c862429102cee03c2ccd338c

Observation 6b9a8a68-1aca-4181-a66f-ff7e7e52d9f7 · outbound

This paper cites LongVU: Spatiotemporal Adaptive Compression for Long Video-Language Understanding.

STAC: Selective Spatiotemporal Aggregation and Compression for Video Reasoning Segmentation LongVU: Spatiotemporal Adaptive Compression for Long Video-Language Understanding

Reference 66

Resolution
unresolved
no resolver link, observed 2026-07-12T06:06:47.233814Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T06:06:47.233814Z digest=sha256:35ff0da67bdce32a079d9601000a2471480d32bb319edb2d22c3c57d0c1936d3

Observation b695f4fe-31ef-4978-bec3-337a0f1036c8 · outbound

This paper cites In: CVPR.

STAC: Selective Spatiotemporal Aggregation and Compression for Video Reasoning Segmentation In: CVPR

Reference 67

Resolution
unresolved
no resolver link, observed 2026-07-12T06:06:47.233814Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T06:06:47.233814Z digest=sha256:405de065fdab41082b83a7a8429f778b4bbe80844f23fd63a29d962ffdaee37c

Observation c2dab446-ef64-436f-a05f-c75209fd2426 · outbound

This paper cites arXiv preprint arXiv:2508.04369 (2025).

STAC: Selective Spatiotemporal Aggregation and Compression for Video Reasoning Segmentation arXiv preprint arXiv:2508.04369 (2025)

Reference 68

Resolution
unresolved
no resolver link, observed 2026-07-12T06:06:47.233814Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T06:06:47.233814Z digest=sha256:2c0b4c88c4b9689cdeadc4c779b80f7059be8f2fff1c6084521e84a493058e8d

Observation dac7d91b-fe0b-4f15-b4be-493fd14a03ee · outbound

This paper cites In: ICCV.

STAC: Selective Spatiotemporal Aggregation and Compression for Video Reasoning Segmentation In: ICCV

Reference 69

Resolution
unresolved
no resolver link, observed 2026-07-12T06:06:47.233814Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T06:06:47.233814Z digest=sha256:d4ba7cdc465e3a0fac5f3d7c1645c2f57110d6bb192d1c0d34ea346c2a48b011

Observation 7b85d7d3-ee20-421a-84cf-f9c4bade943e · outbound

This paper cites Gemini: A Family of Highly Capable Multimodal Models.

STAC: Selective Spatiotemporal Aggregation and Compression for Video Reasoning Segmentation Gemini: A Family of Highly Capable Multimodal Models

Reference 70

Resolution
unresolved
no resolver link, observed 2026-07-12T06:06:47.233814Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T06:06:47.233814Z digest=sha256:d34e397d9f7efe383c693ffed3ee0d0faece613680d9d6a375c95464fd807e3b

Observation 1a80582f-508d-4f15-9231-5a4d7d6deb9c · outbound

This paper cites LLaMA: Open and Efficient Foundation Language Models.

STAC: Selective Spatiotemporal Aggregation and Compression for Video Reasoning Segmentation LLaMA: Open and Efficient Foundation Language Models

Reference 71

Resolution
unresolved
no resolver link, observed 2026-07-12T06:06:47.233814Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T06:06:47.233814Z digest=sha256:a71dc21d4749c9227472807371c2f6547559814749c901dbd9f7f79d3710dcfc

Observation cbb240cc-5e93-47cc-9801-319ebd9e1b52 · outbound

This paper cites Llama 2: Open Foundation and Fine-Tuned Chat Models.

STAC: Selective Spatiotemporal Aggregation and Compression for Video Reasoning Segmentation Llama 2: Open Foundation and Fine-Tuned Chat Models

Reference 72

Resolution
unresolved
no resolver link, observed 2026-07-12T06:06:47.233814Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T06:06:47.233814Z digest=sha256:ffad15f3f6b701310b59beba63788abaf2cb08ffc3ed97c284c6b4a406cb2fd1

Observation 7189ef69-2297-4e59-a673-1eb616517dc2 · outbound

This paper cites In: European Conference on Computer Vision.

STAC: Selective Spatiotemporal Aggregation and Compression for Video Reasoning Segmentation In: European Conference on Computer Vision

Reference 73

Resolution
unresolved
no resolver link, observed 2026-07-12T06:06:47.233814Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T06:06:47.233814Z digest=sha256:30056d6c136ee557f095db4956f98b5725bba05418ca7228db4b31656670d222

Observation 73e56558-2548-4938-968d-0b83975b64fc · outbound

This paper cites Machine Intelligence Research (2026).

STAC: Selective Spatiotemporal Aggregation and Compression for Video Reasoning Segmentation Machine Intelligence Research (2026)

Reference 74

Resolution
unresolved
no resolver link, observed 2026-07-12T06:06:47.233814Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T06:06:47.233814Z digest=sha256:228a65789bc1a4c5b34475e721b8f53bba62ede58ad980bf9150cc304a41cf2f

Observation c4f5f634-8189-4261-afb8-f4f95487a011 · outbound

This paper cites In: ICCV.

STAC: Selective Spatiotemporal Aggregation and Compression for Video Reasoning Segmentation In: ICCV

Reference 75

Resolution
unresolved
no resolver link, observed 2026-07-12T06:06:47.233814Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T06:06:47.233814Z digest=sha256:39c9a7759d72208971a049f64e82d53da5436067f982962e3b3968478d9b4304

Observation b8a6e547-2ca3-4297-a687-b5c140103a6d · outbound

This paper cites In: CVPR.

STAC: Selective Spatiotemporal Aggregation and Compression for Video Reasoning Segmentation In: CVPR

Reference 76

Resolution
unresolved
no resolver link, observed 2026-07-12T06:06:47.233814Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T06:06:47.233814Z digest=sha256:ee60aadb1747edd91c57b2a51cdd47355b87c1ba4aa0ec3c83bb7fae1c2243b5

Observation ba25135d-c025-4de1-b363-f149987afac8 · outbound

This paper cites YouTube-VOS: A Large-Scale Video Object Segmentation Benchmark.

STAC: Selective Spatiotemporal Aggregation and Compression for Video Reasoning Segmentation YouTube-VOS: A Large-Scale Video Object Segmentation Benchmark

Reference 77

Resolution
unresolved
no resolver link, observed 2026-07-12T06:06:47.233814Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T06:06:47.233814Z digest=sha256:1ea47c4a24d0e58e7dadcfa00137b797b3cbfb3112c4f23cd1655e71f5bd9602

Observation b10c7c05-e02b-4bea-8777-3fdf15331f2c · outbound

This paper cites In: ECCV.

STAC: Selective Spatiotemporal Aggregation and Compression for Video Reasoning Segmentation In: ECCV

Reference 78

Resolution
unresolved
no resolver link, observed 2026-07-12T06:06:47.233814Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T06:06:47.233814Z digest=sha256:ae586561cf863d6ba18c851ccb6528efc07c99205d60bb4b24accdfd92a9b094

Observation 7ed5a5f4-7b5f-4753-8be1-b4a8a65b32ec · outbound

This paper cites Vivim: a Video Vision Mamba for Medical Video Segmentation.

STAC: Selective Spatiotemporal Aggregation and Compression for Video Reasoning Segmentation Vivim: a Video Vision Mamba for Medical Video Segmentation

Reference 79

Resolution
unresolved
no resolver link, observed 2026-07-12T06:06:47.233814Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T06:06:47.233814Z digest=sha256:88423929d8436b453a9fd67e545a5bd2e993e49ea0cdd15a4231004611f34b10

Observation 290f63e8-2ef4-48ad-aabc-29890a19ab04 · outbound

This paper cites In: CVPR.

STAC: Selective Spatiotemporal Aggregation and Compression for Video Reasoning Segmentation In: CVPR

Reference 80

Resolution
unresolved
no resolver link, observed 2026-07-12T06:06:47.233814Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T06:06:47.233814Z digest=sha256:cb9323623d2a66e4ae6b3540e2c3669dbeb05a8a7ec4f6610354d7105167a780

Observation f05d300c-997b-450d-ab5b-ea4fbd2a9409 · outbound

This paper cites Sa2VA: Marrying SAM2 with LLaVA for Dense Grounded Understanding of Images and Videos.

STAC: Selective Spatiotemporal Aggregation and Compression for Video Reasoning Segmentation Sa2VA: Marrying SAM2 with LLaVA for Dense Grounded Understanding of Images and Videos

Reference 81

Resolution
unresolved
no resolver link, observed 2026-07-12T06:06:47.233814Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T06:06:47.233814Z digest=sha256:8ad70d8e1b7870beec5625474c3fa12cd4ff00c2368b9b01d981ed59f2ceca6d

Observation 33889bac-850f-4aeb-bee4-9b6d824e4b2e · outbound

This paper cites In: ICCV.

STAC: Selective Spatiotemporal Aggregation and Compression for Video Reasoning Segmentation In: ICCV

Reference 82

Resolution
unresolved
no resolver link, observed 2026-07-12T06:06:47.233814Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T06:06:47.233814Z digest=sha256:94c741c4fd0ca61c1f41df538c54b3305f6408a60790170325fc6bf1d7585a4f

Observation 8989a9ea-f1ed-4eb1-8c16-26734c29cd64 · outbound

This paper cites Video-LLaMA: An Instruction-tuned Audio-Visual Language Model for Video Understanding.

STAC: Selective Spatiotemporal Aggregation and Compression for Video Reasoning Segmentation Video-LLaMA: An Instruction-tuned Audio-Visual Language Model for Video Understanding

Reference 83

Resolution
unresolved
no resolver link, observed 2026-07-12T06:06:47.233814Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T06:06:47.233814Z digest=sha256:34205aa2fc96a844971d3dfa97e6ab0a6cf723edb05ff26fb6fdaf1cdf74276d

Observation e369217a-d568-4ef6-a4b4-d63108ace0bb · outbound

This paper cites In: AAAI.

STAC: Selective Spatiotemporal Aggregation and Compression for Video Reasoning Segmentation In: AAAI

Reference 84

Resolution
unresolved
no resolver link, observed 2026-07-12T06:06:47.233814Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T06:06:47.233814Z digest=sha256:bcd037a01479118fe45b940f1d2b7c741f67032702a57f6dec400ce2593df152

Observation 8314b72e-9a63-4657-acf5-cebd372b1211 · outbound

This paper cites Tracking with Human-Intent Reasoning.

STAC: Selective Spatiotemporal Aggregation and Compression for Video Reasoning Segmentation Tracking with Human-Intent Reasoning

Reference 85

Resolution
unresolved
no resolver link, observed 2026-07-12T06:06:47.233814Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T06:06:47.233814Z digest=sha256:008ecc44db78de4100ab46279079fadf20f44163f01ceed10bf8d4eb35559fed

Observation 796b549c-8f3b-487a-87dc-907d78ce2ac4 · outbound

This paper cites an unresolved cited work.

STAC: Selective Spatiotemporal Aggregation and Compression for Video Reasoning Segmentation Unresolved cited work

Reference 86

Resolution
unresolved
no resolver link, observed 2026-07-12T06:06:47.233814Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T06:06:47.233814Z digest=sha256:e5a2b79eace7d662ea5ab46fd178908005f607854f87315513f0e2280bb4c261

Pith citing papers

No inbound Pith citation observations are available.