Pith. sign in

Paper Citation Record · LEDGER

MPG-SAM 2: Adapting SAM 2 with Mask Priors and Global Context for Referring Video Object Segmentation

As of 11 August 2026, this Paper Citation Record lists 61 of 61 outbound references and 0 inbound Pith citation observations for arXiv:2501.13667.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2501.13667 v5

Coverage vector

measured 61 of 61 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-10T15:47:59.927188Z

measured 61 of 61 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

61 of 61 outbound references displayed

  • verified exact1
  • verified fuzzy38
  • unresolved22
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 4e1c3f6c-bbc6-4ba5-8b13-9314fc777a8e · outbound

This paper cites Xmem++: Production-level video segmentation from few annotated frames.

MPG-SAM 2: Adapting SAM 2 with Mask Priors and Global Context for Referring Video Object Segmentation Xmem++: Production-level video segmentation from few annotated frames

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T15:48:01.036615Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T15:47:59.638768Z digest=sha256:087582362bd65251f3ae19b39f31b8f804f0fdd292e3e9d628786083f9051c1d

Observation bbc84f67-177f-4a4d-9f71-db45046f149d · outbound

This paper cites RefVOS: A Closer Look at Referring Expressions for Video Object Segmentation.

MPG-SAM 2: Adapting SAM 2 with Mask Priors and Global Context for Referring Video Object Segmentation RefVOS: A Closer Look at Referring Expressions for Video Object Segmentation

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-10T15:47:59.644137Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T15:47:59.644137Z digest=sha256:7f892a1a22db8799d70eae782db9bc52bf671775bd22dd4a74b897eacfb323f6

Observation 6bf26444-f89a-4366-b158-ec586a8a771a · outbound

This paper cites End-to-end referring video object segmentation with multi- modal transformers.

MPG-SAM 2: Adapting SAM 2 with Mask Priors and Global Context for Referring Video Object Segmentation End-to-end referring video object segmentation with multi- modal transformers

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T15:48:01.019502Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T15:47:59.649698Z digest=sha256:04bd5278650722b8c587697d333fa92b60281ea1e685f46ee29df83a9782548e

Observation fdc8ee81-920a-4426-9562-f90f532b5172 · outbound

This paper cites End-to-end referring video object segmentation with multi- modal transformers.

MPG-SAM 2: Adapting SAM 2 with Mask Priors and Global Context for Referring Video Object Segmentation End-to-end referring video object segmentation with multi- modal transformers

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T15:48:01.001502Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T15:47:59.654829Z digest=sha256:850edce813d92b3f9f497477785dd6bf1a4b8313e4730646f1db7e6a0ee90a0f

Observation 987ce447-7a12-4674-bc10-519771e68145 · outbound

This paper cites End- to-end object detection with transformers.

MPG-SAM 2: Adapting SAM 2 with Mask Priors and Global Context for Referring Video Object Segmentation End- to-end object detection with transformers

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T15:48:00.984032Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T15:47:59.659696Z digest=sha256:d0362bd557a04df128ea69f12bc507584c04fb8f2a6b3278c026c7675135a411

Observation 755fb4b8-5a47-4de1-9e55-e8b1463f784c · outbound

This paper cites Xmem: Long- term video object segmentation with an atkinson-shiffrin memory model.

MPG-SAM 2: Adapting SAM 2 with Mask Priors and Global Context for Referring Video Object Segmentation Xmem: Long- term video object segmentation with an atkinson-shiffrin memory model

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T15:48:00.966930Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T15:47:59.664751Z digest=sha256:e8e697fbecf1d0fc99101ba1f5b57d5839144e33babd830b48dd1e0366cfa266

Observation ff279dd1-d879-4b85-adb0-7c78be7d3270 · outbound

This paper cites Putting the object back into video object segmentation.

MPG-SAM 2: Adapting SAM 2 with Mask Priors and Global Context for Referring Video Object Segmentation Putting the object back into video object segmentation

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T15:48:00.950457Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T15:47:59.670130Z digest=sha256:1350fcc422268332813c7c286a545cb044a41ca349eecbc36924844e8f1e4315

Observation 4546e337-6ea6-4a03-8f0c-ec6d3e3a81d0 · outbound

This paper cites Segment and Track Anything.

MPG-SAM 2: Adapting SAM 2 with Mask Priors and Global Context for Referring Video Object Segmentation Segment and Track Anything

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-10T15:47:59.674573Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T15:47:59.674573Z digest=sha256:28481cae3ec9869469429f93c985268eb17e4b9374aedf7eee40f6cf38ddc643

Observation 055a0632-966b-491b-95e1-cc2a2fbb4493 · outbound

This paper cites Unsupervised Cross-lingual Representation Learning at Scale.

MPG-SAM 2: Adapting SAM 2 with Mask Priors and Global Context for Referring Video Object Segmentation Unsupervised Cross-lingual Representation Learning at Scale

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-10T15:47:59.679455Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T15:47:59.679455Z digest=sha256:715ab286487a9b59b19dc6a900d4cec771af70475d325d55f41885d0630c4786

Observation ff018f14-f355-42b4-b98e-04a998bfec9a · outbound

This paper cites Vision-language transformer and query generation for refer- ring segmentation.

MPG-SAM 2: Adapting SAM 2 with Mask Priors and Global Context for Referring Video Object Segmentation Vision-language transformer and query generation for refer- ring segmentation

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T15:48:00.932309Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T15:47:59.684899Z digest=sha256:1c884e49e28060f2ee8befe8bccb98344a8d96e09c64124f9cb32ddaf14566b6

Observation 0fa86d86-2c75-4729-a592-455d0adf2496 · outbound

This paper cites Mevis: A large-scale benchmark for video segmentation with motion expressions.

MPG-SAM 2: Adapting SAM 2 with Mask Priors and Global Context for Referring Video Object Segmentation Mevis: A large-scale benchmark for video segmentation with motion expressions

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T15:48:00.913517Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T15:47:59.690233Z digest=sha256:b07e2259c6c2b74dcf8b24e07f2a6da1afb57f292d5569f6b028fbe8389fbfb1

Observation 5971daea-c441-47f0-b780-85b500c13e4d · outbound

This paper cites Language-bridged spatial-temporal interaction for referring video object segmentation.

MPG-SAM 2: Adapting SAM 2 with Mask Priors and Global Context for Referring Video Object Segmentation Language-bridged spatial-temporal interaction for referring video object segmentation

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T15:48:00.896976Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T15:47:59.695176Z digest=sha256:067700c992069a2fa2595563d3da7fbe8841632bc374a0a57ee2ca9a9feea1b0

Observation f94ac849-ef94-41d6-ab2c-b0ef10281cb4 · outbound

This paper cites Unified embedding alignment for open-vocabulary video instance segmentation.

MPG-SAM 2: Adapting SAM 2 with Mask Priors and Global Context for Referring Video Object Segmentation Unified embedding alignment for open-vocabulary video instance segmentation

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T15:48:00.881133Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T15:47:59.700026Z digest=sha256:04724631e57d478cbb2ad22c597373457d094a71940aa773c10ff3fa79f0720e

Observation 64500324-77b2-49a1-b524-4c5921b9300e · outbound

This paper cites Html: Hybrid temporal-scale mul- timodal learning framework for referring video object seg- mentation.

MPG-SAM 2: Adapting SAM 2 with Mask Priors and Global Context for Referring Video Object Segmentation Html: Hybrid temporal-scale mul- timodal learning framework for referring video object seg- mentation

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T15:48:00.865317Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T15:47:59.704741Z digest=sha256:ae82eec54abdaf8e0f00b6837a0eb749b0105aed00c1e4023b77ea5ef50e95b1

Observation 160ea41c-b019-473c-9647-601d3cdd12c5 · outbound

This paper cites Decoupling static and hier- archical motion perception for referring video segmentation.

MPG-SAM 2: Adapting SAM 2 with Mask Priors and Global Context for Referring Video Object Segmentation Decoupling static and hier- archical motion perception for referring video segmentation

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T15:48:00.848532Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T15:47:59.709388Z digest=sha256:3324045dc2983d3575ae5b6fedd79f14c3da9fa592d7fb3481b995e8d5cf2eb6

Observation 7161f8c8-2472-4e51-8763-388d1fb61b66 · outbound

This paper cites Unleashing the Temporal-Spatial Reasoning Capacity of GPT for Training-Free Audio and Language Referenced Video Object Segmentation.

MPG-SAM 2: Adapting SAM 2 with Mask Priors and Global Context for Referring Video Object Segmentation Unleashing the Temporal-Spatial Reasoning Capacity of GPT for Training-Free Audio and Language Referenced Video Object Segmentation

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-10T15:47:59.714033Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T15:47:59.714033Z digest=sha256:37308340138ea81a2769021165f5f00d6e5bb1fb06e44581ca8abb5ee8f76df8

Observation 69c6c784-e421-4ba8-806c-7309be4bb745 · outbound

This paper cites Segment anything in high qual- ity.Advances in Neural Information Processing Systems, 36,.

MPG-SAM 2: Adapting SAM 2 with Mask Priors and Global Context for Referring Video Object Segmentation Segment anything in high qual- ity.Advances in Neural Information Processing Systems, 36,

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-10T15:47:59.718988Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T15:47:59.718988Z digest=sha256:698f2afab9892d408b55c2f144878827a17ab589f8601cbefbacb4bd36882980

Observation dfe47324-1f8c-4d7a-aaad-9288711f0ff0 · outbound

This paper cites Video object segmentation with language referring expressions.

MPG-SAM 2: Adapting SAM 2 with Mask Priors and Global Context for Referring Video Object Segmentation Video object segmentation with language referring expressions

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T15:48:00.815303Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T15:47:59.723933Z digest=sha256:d9ab1ae81565fe17a82e9a759d2f51e530b6660a1268c79b89b8997bc1d44716

Observation 4e3d5cfb-376c-4f68-a950-94455ae29c78 · outbound

This paper cites Segment any- thing.

MPG-SAM 2: Adapting SAM 2 with Mask Priors and Global Context for Referring Video Object Segmentation Segment any- thing

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T15:48:00.797935Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T15:47:59.728435Z digest=sha256:3cdd0a5550faa4b5cf8124426ac3f0ebaccd1a2b5badf85f2687f2d018eeb04c

Observation b89bf872-4ce4-4000-89c9-2bd0ba6b7222 · outbound

This paper cites Lisa: Reasoning segmentation via large language model.

MPG-SAM 2: Adapting SAM 2 with Mask Priors and Global Context for Referring Video Object Segmentation Lisa: Reasoning segmentation via large language model

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T15:48:00.780736Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T15:47:59.733316Z digest=sha256:6ea6e37ec9dce6972f73db74d398c212cd538fa11f5a21cd0434d9b0b2cf7cbb

Observation ced4a7ee-c08d-48ca-af59-a115b9104a63 · outbound

This paper cites Learning to learn better for video object segmentation.

MPG-SAM 2: Adapting SAM 2 with Mask Priors and Global Context for Referring Video Object Segmentation Learning to learn better for video object segmentation

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T15:48:00.764725Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T15:47:59.737900Z digest=sha256:e0ba2e78c60a3cfd0d8c0d3377e9f450ae506aff34be2662646cec85329fab51

Observation 328add4e-aec4-408f-9b64-772b9545be10 · outbound

This paper cites an unresolved cited work.

MPG-SAM 2: Adapting SAM 2 with Mask Priors and Global Context for Referring Video Object Segmentation Unresolved cited work

Reference 22

Resolution
unresolved
raw_fallback, observed 2026-08-10T15:48:00.749283Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T15:47:59.742330Z digest=sha256:bdbc0de13912479327daa40c863e75be3cdc9c6fd8d5b6b01230f9da8e0c4ec3

Observation 5b791fd6-2d71-4e4c-a4a1-6985029cb2a2 · outbound

This paper cites Bidirectional correlation-driven inter-frame inter- action transformer for referring video object segmentation.

MPG-SAM 2: Adapting SAM 2 with Mask Priors and Global Context for Referring Video Object Segmentation Bidirectional correlation-driven inter-frame inter- action transformer for referring video object segmentation

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T15:48:00.733566Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T15:47:59.747025Z digest=sha256:bd7d0ee76ce21bd9af97493235bbfe89db106b20cbca85cf9bf888f743849d96

Observation 9701dfa4-8e8e-4eef-886e-cd73a0ea8d1c · outbound

This paper cites You only infer once: Cross-modal meta-transfer for referring video object segmentation.

MPG-SAM 2: Adapting SAM 2 with Mask Priors and Global Context for Referring Video Object Segmentation You only infer once: Cross-modal meta-transfer for referring video object segmentation

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T15:48:00.717969Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T15:47:59.751536Z digest=sha256:3df8a7a27610095c7af84d77445a1101d61fa0aa3b8ac301a9d8656af58cdbd0

Observation e5c6324e-ed1f-44bc-9d23-12fb89929b46 · outbound

This paper cites RefSAM: Efficiently Adapting Segmenting Anything Model for Referring Video Object Segmentation.

MPG-SAM 2: Adapting SAM 2 with Mask Priors and Global Context for Referring Video Object Segmentation RefSAM: Efficiently Adapting Segmenting Anything Model for Referring Video Object Segmentation

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-10T15:47:59.756198Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T15:47:59.756198Z digest=sha256:e016022594483e24472a82f2f1495cbb20aba05c4b44535be024f561370526f0

Observation 3e9f7077-2c38-4e6e-99f5-159afcd2d098 · outbound

This paper cites Visual instruction tuning.Advances in neural information processing systems, 36, 2024.

MPG-SAM 2: Adapting SAM 2 with Mask Priors and Global Context for Referring Video Object Segmentation Visual instruction tuning.Advances in neural information processing systems, 36, 2024

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-10T15:47:59.761037Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T15:47:59.761037Z digest=sha256:31532c7c3b63c1610e7fae2e09a6d3882b65ce6c1408d32e6eff41658974d3d9

Observation fd5bc5ec-19e1-40f6-b456-e6f50d648946 · outbound

This paper cites Grounding DINO: Marrying DINO with Grounded Pre-Training for Open-Set Object Detection.

MPG-SAM 2: Adapting SAM 2 with Mask Priors and Global Context for Referring Video Object Segmentation Grounding DINO: Marrying DINO with Grounded Pre-Training for Open-Set Object Detection

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-10T15:47:59.765578Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T15:47:59.765578Z digest=sha256:2e620a319c412bd4e03be9d448e00f769aa28ca547135c2c13d2bad44a9a848e

Observation 230a054c-9723-4a4a-93ea-8bf4aac4c325 · outbound

This paper cites Decoupled weight decay regularization.

MPG-SAM 2: Adapting SAM 2 with Mask Priors and Global Context for Referring Video Object Segmentation Decoupled weight decay regularization

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-10T15:47:59.770551Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T15:47:59.770551Z digest=sha256:cb7b8c248aea17c6a6dcb230c96ee1d35521654fad43308c6d6641d53f4dfbae

Observation d947f84c-2012-4c4a-94d4-3e539f1415bd · outbound

This paper cites Soc: Semantic-assisted object cluster for referring video object segmentation.Advances in Neural Information Processing Systems, 36, 2024.

MPG-SAM 2: Adapting SAM 2 with Mask Priors and Global Context for Referring Video Object Segmentation Soc: Semantic-assisted object cluster for referring video object segmentation.Advances in Neural Information Processing Systems, 36, 2024

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T15:48:00.682289Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T15:47:59.775051Z digest=sha256:d6fea0dbd44ef4812be71beef0e81245d17f3e5da05ba37ff92fe62adb9beff0

Observation 50a38455-8280-4a0f-9131-321f32e3234f · outbound

This paper cites Generation and comprehension of unambiguous object descriptions.

MPG-SAM 2: Adapting SAM 2 with Mask Priors and Global Context for Referring Video Object Segmentation Generation and comprehension of unambiguous object descriptions

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-10T15:47:59.779445Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T15:47:59.779445Z digest=sha256:83d896ced3108c116b60babfe3fc70106a935dc2b5d953f3bac1e0b34ba9c917

Observation bd1b8c88-a717-449d-a81c-135e74c8e4ab · outbound

This paper cites Visual-textual capsule routing for text-based video segmentation.

MPG-SAM 2: Adapting SAM 2 with Mask Priors and Global Context for Referring Video Object Segmentation Visual-textual capsule routing for text-based video segmentation

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T15:48:00.655677Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T15:47:59.784047Z digest=sha256:f1b853cbc34069a4caaccaefe0ecdd3a44f8dbfd2b1ce968aa917bc6b5788516

Observation 38e25771-d08d-407a-a24a-19c3b8c25675 · outbound

This paper cites Spectrum-guided multi-granularity referring video object segmentation.

MPG-SAM 2: Adapting SAM 2 with Mask Priors and Global Context for Referring Video Object Segmentation Spectrum-guided multi-granularity referring video object segmentation

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T15:48:00.640795Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T15:47:59.788681Z digest=sha256:4f875256d640683e7a8357766ff2c7dbab04de2ec91c5b20af0727eb4bf670ab

Observation 04df44cf-137c-43d2-9023-03e4e5a2394f · outbound

This paper cites V-net: Fully convolutional neural networks for volumetric medical image segmentation.

MPG-SAM 2: Adapting SAM 2 with Mask Priors and Global Context for Referring Video Object Segmentation V-net: Fully convolutional neural networks for volumetric medical image segmentation

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T15:48:00.624776Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T15:47:59.793025Z digest=sha256:9bc572230ee2beb8270547ccc6f02032bdf7d1d06b29fe359e11314ed38eaf01

Observation ad7d8cdc-6449-4416-bddc-15badaeee255 · outbound

This paper cites Video object segmentation using space-time memory networks.

MPG-SAM 2: Adapting SAM 2 with Mask Priors and Global Context for Referring Video Object Segmentation Video object segmentation using space-time memory networks

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T15:48:00.609152Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T15:47:59.797817Z digest=sha256:4fe5b244364b3297ffb8b2e6c0f7a47b3026695b7b9ab0141ae5012f46329614

Observation b476880e-13e2-4bc3-ba23-53d3216bf272 · outbound

This paper cites Semantic and sequential alignment for referring video object segmentation.

MPG-SAM 2: Adapting SAM 2 with Mask Priors and Global Context for Referring Video Object Segmentation Semantic and sequential alignment for referring video object segmentation

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T15:48:00.593023Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T15:47:59.802324Z digest=sha256:5c4b9544b42999ee821b3d67ac92f918cfacc796886be65b35104de2af5b9d1b

Observation a2d4b4e7-3df5-46c6-88d9-7eb0fb05c1c2 · outbound

This paper cites The 2017 DAVIS Challenge on Video Object Segmentation.

MPG-SAM 2: Adapting SAM 2 with Mask Priors and Global Context for Referring Video Object Segmentation The 2017 DAVIS Challenge on Video Object Segmentation

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-10T15:47:59.807087Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T15:47:59.807087Z digest=sha256:132347d6f06a85ace1f5a803cfcd01886684c90e995445b6f8e4f9c884157cff

Observation e384604e-d98d-4289-b000-5cd9177d1b1e · outbound

This paper cites Glamm: Pixel grounding large multimodal model.

MPG-SAM 2: Adapting SAM 2 with Mask Priors and Global Context for Referring Video Object Segmentation Glamm: Pixel grounding large multimodal model

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-10T15:47:59.811931Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T15:47:59.811931Z digest=sha256:e134ff649e5c7f868ea8dd2c7b21b7f5634675df17526761affd301549e251c7

Observation 36430e9f-d503-4473-b3e4-ed7f044aac35 · outbound

This paper cites SAM 2: Segment Anything in Images and Videos.

MPG-SAM 2: Adapting SAM 2 with Mask Priors and Global Context for Referring Video Object Segmentation SAM 2: Segment Anything in Images and Videos

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-10T15:47:59.817076Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T15:47:59.817076Z digest=sha256:edeb88ca7185b322e235fa367a5d931d184d3f3a183331f490db3999c88cdfe2

Observation 1dd84681-af88-483f-aac6-521cc1ddbc2e · outbound

This paper cites Cus- tomized sam 2 for referring remote sensing image segmenta- tion.arXiv preprint arXiv:2503.07266, 2025.

MPG-SAM 2: Adapting SAM 2 with Mask Priors and Global Context for Referring Video Object Segmentation Cus- tomized sam 2 for referring remote sensing image segmenta- tion.arXiv preprint arXiv:2503.07266, 2025

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-10T15:47:59.821941Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T15:47:59.821941Z digest=sha256:dc7300a1fc14403ce8052951a82695cadd70ec86f3682489708446382cdb86b8

Observation 92ed7359-2cfb-4b8d-985e-b3e434a333e0 · outbound

This paper cites Urvos: Unified referring video object segmentation network with a large-scale benchmark.

MPG-SAM 2: Adapting SAM 2 with Mask Priors and Global Context for Referring Video Object Segmentation Urvos: Unified referring video object segmentation network with a large-scale benchmark

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T15:48:00.567693Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T15:47:59.826708Z digest=sha256:89a7c0545f87fe2e19696a9be3f62fffbbcc477f10cc516d44abcbc16b2e7f55

Observation 94e8d278-d03b-4049-8de5-aca91dbf4e0b · outbound

This paper cites Temporal collection and distribution for referring video object segmentation.

MPG-SAM 2: Adapting SAM 2 with Mask Priors and Global Context for Referring Video Object Segmentation Temporal collection and distribution for referring video object segmentation

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T15:48:00.552878Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T15:47:59.831198Z digest=sha256:c05b410acad64dc2c84fa4ac1146e15b859fb1298e1c2dadf32cae069dc9f257

Observation 3065cf5b-7afd-4526-833f-13bb09643fd2 · outbound

This paper cites Samrs: Scaling-up re- mote sensing segmentation dataset with segment anything model.Advances in Neural Information Processing Systems, 36, 2024.

MPG-SAM 2: Adapting SAM 2 with Mask Priors and Global Context for Referring Video Object Segmentation Samrs: Scaling-up re- mote sensing segmentation dataset with segment anything model.Advances in Neural Information Processing Systems, 36, 2024

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T15:48:00.537906Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T15:47:59.835790Z digest=sha256:d1085352a8952553ddd0aedac0d284da8e328180faacf69bad93c799136035fa

Observation a1d6c23c-9577-466b-b38a-327dcc6a4f30 · outbound

This paper cites Asymmetric cross-guided attention network for actor and ac- tion video segmentation from natural language query.

MPG-SAM 2: Adapting SAM 2 with Mask Priors and Global Context for Referring Video Object Segmentation Asymmetric cross-guided attention network for actor and ac- tion video segmentation from natural language query

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T15:48:00.521970Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T15:47:59.840270Z digest=sha256:14bf3d4c3c14ca03b43fed3b90ba607c0afaf1c5440c4ca41cef45e5dd27ece1

Observation 5959649d-ce2c-46b4-9e05-e8c3ca604f26 · outbound

This paper cites Image as a foreign language: Beit pretraining for vision and vision- language tasks.

MPG-SAM 2: Adapting SAM 2 with Mask Priors and Global Context for Referring Video Object Segmentation Image as a foreign language: Beit pretraining for vision and vision- language tasks

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T15:48:00.504760Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T15:47:59.845067Z digest=sha256:0e18afd8ee568c0ca1d0f7ab88c402bddda0f92e3fec9295e9528a19124dd659

Observation 8f1e15ea-a785-4cc4-ace7-b3b1e43c0986 · outbound

This paper cites HyperSeg: Towards Universal Visual Segmentation with Large Language Model.

MPG-SAM 2: Adapting SAM 2 with Mask Priors and Global Context for Referring Video Object Segmentation HyperSeg: Towards Universal Visual Segmentation with Large Language Model

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-10T15:47:59.849646Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T15:47:59.849646Z digest=sha256:1f6dfb168d8929989f7686b8e0628db354303e58903e7230e71ff42c0d9d6c72

Observation 99a2b93d-69c9-4020-8819-a18dc653a03a · outbound

This paper cites Multi-level representation learning with semantic alignment for referring video object segmentation.

MPG-SAM 2: Adapting SAM 2 with Mask Priors and Global Context for Referring Video Object Segmentation Multi-level representation learning with semantic alignment for referring video object segmentation

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T15:48:00.488432Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T15:47:59.854413Z digest=sha256:4160efead884b7358ccac32739cdb4a30fac831d04269dc0a92051c4e0556c69

Observation dcaacb23-6094-40d8-a4eb-12d4d8299c9d · outbound

This paper cites Onlinerefer: A simple online baseline for referring video object segmentation.

MPG-SAM 2: Adapting SAM 2 with Mask Priors and Global Context for Referring Video Object Segmentation Onlinerefer: A simple online baseline for referring video object segmentation

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T15:48:00.471795Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T15:47:59.858650Z digest=sha256:065a46ebe36651c5127b66975fc75982a72ffa2609be018b963ff41b6bbc6ace

Observation 27dad8b8-5237-45d9-b191-642f246c9a2d · outbound

This paper cites Language as queries for referring video object segmen- tation.

MPG-SAM 2: Adapting SAM 2 with Mask Priors and Global Context for Referring Video Object Segmentation Language as queries for referring video object segmen- tation

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T15:48:00.455585Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T15:47:59.863235Z digest=sha256:8936b62b7b942b6bfdabb88ebcca02591af58471cc5fec4b511bdef6351d5a4b

Observation af69f5b1-cba9-4ea5-bb47-dfc980093eef · outbound

This paper cites Logiczsl: Exploring logic- induced representation for compositional zero-shot learning.

MPG-SAM 2: Adapting SAM 2 with Mask Priors and Global Context for Referring Video Object Segmentation Logiczsl: Exploring logic- induced representation for compositional zero-shot learning

Reference 49

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T15:48:00.438701Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T15:47:59.868031Z digest=sha256:a985d531ee3de20cd7dbd39cf131cca837ce5a4ffdbe7b7ec3d698d4e58a9a99

Observation 3690eaf6-802f-4159-9288-f20a2e6b9847 · outbound

This paper cites Efficientsam: Leveraged masked image pretraining for efficient segment anything.

MPG-SAM 2: Adapting SAM 2 with Mask Priors and Global Context for Referring Video Object Segmentation Efficientsam: Leveraged masked image pretraining for efficient segment anything

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T15:48:00.422489Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T15:47:59.873004Z digest=sha256:617a78b5d1c212c527d2aa862bd167f0990ba5d07e2deadf0567d5d39fd4601c

Observation d52ffb01-c86d-4286-9866-779efd545e76 · outbound

This paper cites u-LLaVA: Unifying Multi-Modal Tasks via Large Language Model.

MPG-SAM 2: Adapting SAM 2 with Mask Priors and Global Context for Referring Video Object Segmentation u-LLaVA: Unifying Multi-Modal Tasks via Large Language Model

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-10T15:47:59.877607Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T15:47:59.877607Z digest=sha256:7dc7a28d8685f8dfaf898a1f49e64e1bdc2ce9eaeb9e49a0ee4763ccc3bb5d62

Observation de3c20d8-e016-42fd-8746-f16c4ab22481 · outbound

This paper cites Visa: Reasoning video object segmentation via large language models.

MPG-SAM 2: Adapting SAM 2 with Mask Priors and Global Context for Referring Video Object Segmentation Visa: Reasoning video object segmentation via large language models

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T15:48:00.405801Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T15:47:59.882837Z digest=sha256:f5b958f2f33c7f1e4b1ebe7f199cf037e10b6b190d1f035d1207e5877b0063d3

Observation f36f6979-76b4-4c7c-9717-e82879d29075 · outbound

This paper cites Referred by multi-modality: A unified tem- poral transformer for video object segmentation.

MPG-SAM 2: Adapting SAM 2 with Mask Priors and Global Context for Referring Video Object Segmentation Referred by multi-modality: A unified tem- poral transformer for video object segmentation

Reference 53

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T15:48:00.390008Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T15:47:59.887488Z digest=sha256:082c0b40153662c2f2add3226f97f59ae493cf639297643e7b350159c7f7b8d2

Observation 39fab523-0cbf-4cde-a3ea-496bc9ffb894 · outbound

This paper cites Modeling context in referring expres- sions.

MPG-SAM 2: Adapting SAM 2 with Mask Priors and Global Context for Referring Video Object Segmentation Modeling context in referring expres- sions

Reference 54

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T15:48:00.372685Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T15:47:59.892095Z digest=sha256:b53fcf0c681eb8ca318b9c515252e628afffe6f2f4f5057fdb7ec8a07871b8da

Observation cc420157-1065-4e3b-a348-dada5e697f3f · outbound

This paper cites A Simple Baseline with Single-encoder for Referring Image Segmentation.

MPG-SAM 2: Adapting SAM 2 with Mask Priors and Global Context for Referring Video Object Segmentation A Simple Baseline with Single-encoder for Referring Image Segmentation

Reference 55

Resolution
verified exact
local_arxiv, observed 2026-08-10T15:48:00.007696Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T15:47:59.897025Z digest=sha256:c99e69d90235e3eb1f0db88863a71f3859a5664661688485c0765baaf08006d3

Observation 959aed02-6d41-47d3-9e66-7763d8836a05 · outbound

This paper cites Losh: Long-short text joint prediction network for referring video object segmentation.

MPG-SAM 2: Adapting SAM 2 with Mask Priors and Global Context for Referring Video Object Segmentation Losh: Long-short text joint prediction network for referring video object segmentation

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-10T15:47:59.901850Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T15:47:59.901850Z digest=sha256:99ad2fecd011cbbf85ed9c22440d9553d28dbdeaa3b99ea7e4abad768e7e613d

Observation 015c654e-2242-41cf-a9d4-161540a8f6cb · outbound

This paper cites Surgicalsam: Efficient class prompt- able surgical instrument segmentation.

MPG-SAM 2: Adapting SAM 2 with Mask Priors and Global Context for Referring Video Object Segmentation Surgicalsam: Efficient class prompt- able surgical instrument segmentation

Reference 57

Resolution
unresolved
no resolver link, observed 2026-08-10T15:47:59.906540Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T15:47:59.906540Z digest=sha256:4350eeebcb07b05caa7a3ca8c035b75e224d9d2d60f9621aed9cc0972ba8c0fe

Observation 42fa24c1-7262-43ff-885e-fd50ec4fb259 · outbound

This paper cites Faster Segment Anything: Towards Lightweight SAM for Mobile Applications.

MPG-SAM 2: Adapting SAM 2 with Mask Priors and Global Context for Referring Video Object Segmentation Faster Segment Anything: Towards Lightweight SAM for Mobile Applications

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-10T15:47:59.911548Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T15:47:59.911548Z digest=sha256:11220cc7409a10f612c0b209771606d12694bde991a69e6c079ca65684806310

Observation 747b87e0-fc36-405d-a4e0-52a21ae4e60d · outbound

This paper cites EVF-SAM: Early Vision-Language Fusion for Text-Prompted Segment Anything Model.

MPG-SAM 2: Adapting SAM 2 with Mask Priors and Global Context for Referring Video Object Segmentation EVF-SAM: Early Vision-Language Fusion for Text-Prompted Segment Anything Model

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-10T15:47:59.916575Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T15:47:59.916575Z digest=sha256:166805236b9611fc0728cc022c64b35fec6a6338b2bbd17e0176cce899cdd347

Observation a569ecdb-9f60-4755-9229-2bedc8ca6b6c · outbound

This paper cites Deformable detr: Deformable transformers for end-to-end object detection.

MPG-SAM 2: Adapting SAM 2 with Mask Priors and Global Context for Referring Video Object Segmentation Deformable detr: Deformable transformers for end-to-end object detection

Reference 60

Resolution
unresolved
no resolver link, observed 2026-08-10T15:47:59.922407Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T15:47:59.922407Z digest=sha256:92114758bc24999ebf24161408439a74318085d2880e6235aa8874796663c20a

Observation 0cd8a6d0-c9f7-4798-b2df-3d797451fb83 · outbound

This paper cites Exploring pre-trained text- to-video diffusion models for referring video object segmen- tation.

MPG-SAM 2: Adapting SAM 2 with Mask Priors and Global Context for Referring Video Object Segmentation Exploring pre-trained text- to-video diffusion models for referring video object segmen- tation

Reference 61

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T15:48:00.326524Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T15:47:59.927188Z digest=sha256:6a5d9ac4f84e1a17816238e2ec8648b9a2634a7d50fc994b048bb4e5e38d43da

Pith citing papers

No inbound Pith citation observations are available.