Pith. sign in

Paper Citation Record · LEDGER

ViewSAM: Learning View-aware Cross-modal Semantics for Weakly Supervised Cross-view Referring Multi-Object Tracking

As of 11 August 2026, this Paper Citation Record lists 43 of 43 outbound references and 0 inbound Pith citation observations for arXiv:2605.02638.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2605.02638 v1

Coverage vector

measured 43 of 43 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-05-08T18:32:38.644948Z

measured 43 of 43 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-11T06:34:44.6726+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

43 of 43 outbound references displayed

  • verified exact6
  • verified fuzzy35
  • unresolved2
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 2a4cd945-208c-44a2-9cdd-2045feef5fff · outbound

This paper cites Cross-view referring multi-object tracking.

ViewSAM: Learning View-aware Cross-modal Semantics for Weakly Supervised Cross-view Referring Multi-Object Tracking Cross-view referring multi-object tracking

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-05-26T04:46:43.391189Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-08T18:32:38.644948Z digest=sha256:406da29dc7a31f25b907660c0b4a505c0f2f545601a6264ce931810b935dc5a1

Observation dcc8dbfb-9ba0-4df4-ab9c-423d0e6e2625 · outbound

This paper cites Referring multi-object tracking.

ViewSAM: Learning View-aware Cross-modal Semantics for Weakly Supervised Cross-view Referring Multi-Object Tracking Referring multi-object tracking

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-05-26T04:46:43.361789Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-08T18:32:38.644948Z digest=sha256:50056fe48eb0a25901170c5fc784cb12a0c4afd61b81b3a3ef2004d86dbe271e

Observation b9f3c2a9-c7b3-4e14-88ac-e00fb53e5496 · outbound

This paper cites CC-3DT: Panoramic 3D Object Tracking via Cross-Camera Fusion.

ViewSAM: Learning View-aware Cross-modal Semantics for Weakly Supervised Cross-view Referring Multi-Object Tracking CC-3DT: Panoramic 3D Object Tracking via Cross-Camera Fusion

Reference 3

Resolution
verified exact
arxiv_id, observed 2026-05-09T06:20:42.583740Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-08T18:32:38.644948Z digest=sha256:46a1592b36a25b2e0fe745045c62156e7c5c34ec8766cfe59e45bf33001abeb1

Observation afa863a4-5ba1-4570-9d5c-fd92b1dcffc5 · outbound

This paper cites Tango: training- free embodied ai agents for open-world tasks.

ViewSAM: Learning View-aware Cross-modal Semantics for Weakly Supervised Cross-view Referring Multi-Object Tracking Tango: training- free embodied ai agents for open-world tasks

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-05-26T04:46:43.374965Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-08T18:32:38.644948Z digest=sha256:56694978586cdbd5efffb37690c7710be8c02d01d654467d6724fad1bf948a7d

Observation f57d16c3-2047-4c11-83c3-37c600aad160 · outbound

This paper cites Divotrack: A novel dataset and baseline method for cross-view multi- object tracking in diverse open scenes.International Journal of Computer Vision, 132(4):1075– 1090.

ViewSAM: Learning View-aware Cross-modal Semantics for Weakly Supervised Cross-view Referring Multi-Object Tracking Divotrack: A novel dataset and baseline method for cross-view multi- object tracking in diverse open scenes.International Journal of Computer Vision, 132(4):1075– 1090

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-05-26T04:46:43.380116Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-08T18:32:38.644948Z digest=sha256:3f454680dab10d754ec742563c7cc3b6115c64efe274558cdba12acdbf9a8362

Observation b0df05a0-2ae8-4cdc-9a42-405be7eb3d42 · outbound

This paper cites Multi-target multi-camera tracking with spatial-temporal network.

ViewSAM: Learning View-aware Cross-modal Semantics for Weakly Supervised Cross-view Referring Multi-Object Tracking Multi-target multi-camera tracking with spatial-temporal network

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-05-26T04:46:43.388665Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-08T18:32:38.644948Z digest=sha256:b22c043004a7c1c19c82bccd43ce7baebdbbdcd002a894d71020302a536554e2

Observation 8514f0bf-3897-4418-a126-c27cf917233c · outbound

This paper cites Dual-head feature enhancement for graph-based cross-view multi-object tracking.

ViewSAM: Learning View-aware Cross-modal Semantics for Weakly Supervised Cross-view Referring Multi-Object Tracking Dual-head feature enhancement for graph-based cross-view multi-object tracking

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-05-26T04:46:43.377590Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-08T18:32:38.644948Z digest=sha256:ec2c7a97a94e00088d904ec7212761e07856a0db5c2de4dcd9391ee9a90c0bad

Observation ff839168-b3ad-4a4f-a14f-3afbccb43d9d · outbound

This paper cites Gmt: Effective global framework for multi-camera multi-target tracking.arXiv e-prints, pages arXiv–2407.

ViewSAM: Learning View-aware Cross-modal Semantics for Weakly Supervised Cross-view Referring Multi-Object Tracking Gmt: Effective global framework for multi-camera multi-target tracking.arXiv e-prints, pages arXiv–2407

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-05-26T04:46:43.368475Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-08T18:32:38.644948Z digest=sha256:34eb0a09216dfff6f766edaddb12b49f5cebbe1c532df39157a7570084217aa6

Observation 22c59ac3-7189-458e-8179-f798064ae16d · outbound

This paper cites All-day multi- camera multi-target tracking.

ViewSAM: Learning View-aware Cross-modal Semantics for Weakly Supervised Cross-view Referring Multi-Object Tracking All-day multi- camera multi-target tracking

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-05-26T04:46:43.385873Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-08T18:32:38.644948Z digest=sha256:863b4042d68c26ee0c81c20f3f8c8da0a5729393ce599e7e080d0b0dfca9e63f

Observation f32527cf-8940-45e7-832f-a6e56dbc4d8a · outbound

This paper cites Consis- tencies are all you need for semi-supervised vision-language tracking.

ViewSAM: Learning View-aware Cross-modal Semantics for Weakly Supervised Cross-view Referring Multi-Object Tracking Consis- tencies are all you need for semi-supervised vision-language tracking

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-05-26T04:46:43.382923Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-08T18:32:38.644948Z digest=sha256:88844cd72f950bb849c18a2cac95d1e511de8496d3240c3551a251f063d5cbf4

Observation bd8eba3f-82f4-4550-9bcc-73db70a063ab · outbound

This paper cites Large-margin weakly supervised dimen- sionality reduction.

ViewSAM: Learning View-aware Cross-modal Semantics for Weakly Supervised Cross-view Referring Multi-Object Tracking Large-margin weakly supervised dimen- sionality reduction

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-05-26T04:46:43.365165Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-08T18:32:38.644948Z digest=sha256:1bbc9c38e2cecac4bfb271bc324f6e236b6f18d7d51c699251dbaddfc32a5467

Observation be6285cf-6a30-417f-9f72-9c89ee55cd4d · outbound

This paper cites Weaksam: Segment anything meets weakly-supervised instance-level recognition.

ViewSAM: Learning View-aware Cross-modal Semantics for Weakly Supervised Cross-view Referring Multi-Object Tracking Weaksam: Segment anything meets weakly-supervised instance-level recognition

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-05-26T04:46:43.347611Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-08T18:32:38.644948Z digest=sha256:86761e4c54e4fa8b06db667e7f23ccb4fbed4fe98bf3c385e856dc77135edb79

Observation 674215f9-e3c8-45d4-a927-0cd24b6dc173 · outbound

This paper cites A brief introduction to weakly supervised learning.National science review, 5(1):44–53.

ViewSAM: Learning View-aware Cross-modal Semantics for Weakly Supervised Cross-view Referring Multi-Object Tracking A brief introduction to weakly supervised learning.National science review, 5(1):44–53

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-05-26T04:46:43.269114Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-08T18:32:38.644948Z digest=sha256:bbdab853962a8a4fb9fdf64fa45358191cdbcd5c6750523a8d0c5109176e370d

Observation a9cdf71e-cad7-427d-b9b2-5b5d5596869b · outbound

This paper cites Language models are few-shot learners.Advances in neural information processing systems, 33:1877–1901.

ViewSAM: Learning View-aware Cross-modal Semantics for Weakly Supervised Cross-view Referring Multi-Object Tracking Language models are few-shot learners.Advances in neural information processing systems, 33:1877–1901

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-05-26T04:46:43.335368Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-08T18:32:38.644948Z digest=sha256:13bc0065e7e57d7c935a7f0423bacb03cd0036e4f8a91b3aa234617889bf3aeb

Observation d97dc494-cc7b-4b51-8422-f8b87bad574b · outbound

This paper cites Learning transferable visual models from natural language supervision.

ViewSAM: Learning View-aware Cross-modal Semantics for Weakly Supervised Cross-view Referring Multi-Object Tracking Learning transferable visual models from natural language supervision

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-05-26T04:46:43.340941Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-08T18:32:38.644948Z digest=sha256:94f45af7624c8b7445a77dda241df7b7e730e8bdbeac2760201d1d2c25d35c1e

Observation d6d07fa7-b19c-416d-9bbf-7e7fcea47958 · outbound

This paper cites Segment anything.

ViewSAM: Learning View-aware Cross-modal Semantics for Weakly Supervised Cross-view Referring Multi-Object Tracking Segment anything

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-05-26T04:46:43.350961Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-08T18:32:38.644948Z digest=sha256:395c83bcab8c4e61dd1079cadb253c9f61b49b0762578a9c76df5af6d9adbe0e

Observation a0b8c718-cafd-4259-b031-2fe85377f9d4 · outbound

This paper cites Sam 2: Segment anything in images and videos.

ViewSAM: Learning View-aware Cross-modal Semantics for Weakly Supervised Cross-view Referring Multi-Object Tracking Sam 2: Segment anything in images and videos

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-05-26T04:46:43.359191Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-08T18:32:38.644948Z digest=sha256:efd07ed187c5c140d0a6e06c1ff791c063fbbdbaabadb6dd36214be844545706

Observation e2dd3cc0-404b-45b0-9791-f8cdf8bdd732 · outbound

This paper cites SAM 3: Segment Anything with Concepts.

ViewSAM: Learning View-aware Cross-modal Semantics for Weakly Supervised Cross-view Referring Multi-Object Tracking SAM 3: Segment Anything with Concepts

Reference 18

Resolution
verified exact
arxiv_id, observed 2026-05-09T06:20:42.591222Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-08T18:32:38.644948Z digest=sha256:d8691a407e90dab691a047ebed0d95158c9dc420b74a87e4156cba96f8e41a8a

Observation c750f0b1-d7b7-4cdb-aebc-652a71247249 · outbound

This paper cites From sam to cams: Exploring segment anything model for weakly supervised semantic segmentation.

ViewSAM: Learning View-aware Cross-modal Semantics for Weakly Supervised Cross-view Referring Multi-Object Tracking From sam to cams: Exploring segment anything model for weakly supervised semantic segmentation

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-05-26T04:46:43.304958Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-08T18:32:38.644948Z digest=sha256:8ee9b3030ca5f14d499cbbb8447e8ff4e6e0624d7858f5474412fb68609de863

Observation 79c744d7-6cd8-4dd4-a9d9-ba5923c8afff · outbound

This paper cites an unresolved cited work.

ViewSAM: Learning View-aware Cross-modal Semantics for Weakly Supervised Cross-view Referring Multi-Object Tracking Unresolved cited work

Reference 20

Resolution
unresolved
raw_fallback, observed 2026-05-26T04:46:43.301794Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-08T18:32:38.644948Z digest=sha256:b873dd2f31dee3d86e982cf2d2a503ea5db379557d647b72c690210df2ffb4af

Observation 565555e0-e1b4-4ed3-aebb-1bdc599d33bb · outbound

This paper cites Tracking by natural language specification.

ViewSAM: Learning View-aware Cross-modal Semantics for Weakly Supervised Cross-view Referring Multi-Object Tracking Tracking by natural language specification

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-05-26T04:46:43.322335Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-08T18:32:38.644948Z digest=sha256:0be039253fbe09ea88b7f105635a0e816fff5f3ddf1d93ec7c78b4039a54c401

Observation 49f2c68a-0f3b-4e5b-89ed-d8cde46001d6 · outbound

This paper cites Towards more flexible and accurate object tracking with natural language: Algorithms and benchmark.

ViewSAM: Learning View-aware Cross-modal Semantics for Weakly Supervised Cross-view Referring Multi-Object Tracking Towards more flexible and accurate object tracking with natural language: Algorithms and benchmark

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-05-26T04:46:43.319233Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-08T18:32:38.644948Z digest=sha256:e455e6104a6f177136bbfafe4238d51c2796f61735e3d783a07100df90cdfd09

Observation 5fb763be-19e5-48fa-8533-e0e90074d693 · outbound

This paper cites Joint visual grounding and tracking with natural language specification.

ViewSAM: Learning View-aware Cross-modal Semantics for Weakly Supervised Cross-view Referring Multi-Object Tracking Joint visual grounding and tracking with natural language specification

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-05-26T04:46:43.288487Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-08T18:32:38.644948Z digest=sha256:971dd0c5062d7feb0928338d5021de817aa8a70c45ea7946c37ccd4e5e809131

Observation 6652e5b5-dcae-479f-bf95-523fe5478ba9 · outbound

This paper cites Divert more attention to vision- language tracking.

ViewSAM: Learning View-aware Cross-modal Semantics for Weakly Supervised Cross-view Referring Multi-Object Tracking Divert more attention to vision- language tracking

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-05-26T04:46:43.331613Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-08T18:32:38.644948Z digest=sha256:319ee77a2cfff4a6fbca67f58de66533617b7d06661123af31f469de7dddb88a

Observation f6a1db19-275c-4d48-b0ce-7a8c05ea5522 · outbound

This paper cites R1-track: Direct application of mllms to visual object tracking via reinforcement learning.

ViewSAM: Learning View-aware Cross-modal Semantics for Weakly Supervised Cross-view Referring Multi-Object Tracking R1-track: Direct application of mllms to visual object tracking via reinforcement learning

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-05-26T04:46:43.325166Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-08T18:32:38.644948Z digest=sha256:63a6020505a5d9ffa39b8daad55b73b67fe6fb69b422ff97b5f2123f941583a4

Observation 488ef27e-1fb6-4e72-b3ea-a2ed0df724df · outbound

This paper cites an unresolved cited work.

ViewSAM: Learning View-aware Cross-modal Semantics for Weakly Supervised Cross-view Referring Multi-Object Tracking Unresolved cited work

Reference 26

Resolution
unresolved
raw_fallback, observed 2026-05-26T04:46:43.328150Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-08T18:32:38.644948Z digest=sha256:792a3b2148fbc1ed02a1ceacff97601be47491fb6cb810574f3626a9c4cccc23

Observation 66f6b1fc-82f4-48b0-9b66-0b3450713deb · outbound

This paper cites ikun: Speak to trackers without retraining.

ViewSAM: Learning View-aware Cross-modal Semantics for Weakly Supervised Cross-view Referring Multi-Object Tracking ikun: Speak to trackers without retraining

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-05-26T04:46:43.298729Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-08T18:32:38.644948Z digest=sha256:ac52f01deddf99ae8ccc1e83737a55779b7ed801cc5149c7df5af108777f365c

Observation 76f9e535-b6d7-49d5-b7e8-d4c0d134c1fc · outbound

This paper cites Lamot: Language-guided multi-object tracking.

ViewSAM: Learning View-aware Cross-modal Semantics for Weakly Supervised Cross-view Referring Multi-Object Tracking Lamot: Language-guided multi-object tracking

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-05-26T04:46:43.295802Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-08T18:32:38.644948Z digest=sha256:bf51cdba622da1adc48d653827625838bfd53e9cd55ad6805e67056568ad3988

Observation 62253a5a-40fc-48f8-894c-a3f5f3a07bab · outbound

This paper cites Language decoupling with fine-grained knowledge guidance for referring multi-object tracking.

ViewSAM: Learning View-aware Cross-modal Semantics for Weakly Supervised Cross-view Referring Multi-Object Tracking Language decoupling with fine-grained knowledge guidance for referring multi-object tracking

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-05-26T04:46:43.308572Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-08T18:32:38.644948Z digest=sha256:f7450493e59c3b8698855a8f66c3ed6564b73ba8b63b74281aab2d7cb2da2b9a

Observation c205ed74-41db-4ea2-9ce3-982e227f1090 · outbound

This paper cites Temporal-enhanced multimodal transformer for referring multi-object tracking and segmentation.IEEE Transactions on Circuits and Systems for Video Technology.

ViewSAM: Learning View-aware Cross-modal Semantics for Weakly Supervised Cross-view Referring Multi-Object Tracking Temporal-enhanced multimodal transformer for referring multi-object tracking and segmentation.IEEE Transactions on Circuits and Systems for Video Technology

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-05-26T04:46:43.292306Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-08T18:32:38.644948Z digest=sha256:897afeb3e66762d682413cb7627d4255026acf4884aecaa623920406dc7daeab

Observation 759eae0c-2d5a-4f99-870e-1b4fab43fcad · outbound

This paper cites Cgatracker: Correlation-aware graph alignment for referring multi-object tracking.IEEE Transactions on Circuits and Systems for Video Technology.

ViewSAM: Learning View-aware Cross-modal Semantics for Weakly Supervised Cross-view Referring Multi-Object Tracking Cgatracker: Correlation-aware graph alignment for referring multi-object tracking.IEEE Transactions on Circuits and Systems for Video Technology

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-05-26T04:46:43.355915Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-08T18:32:38.644948Z digest=sha256:19879f533d252d723dbe14d1420e5146b6fd5b3c254b38e0410dd6317cce532b

Observation 109fce50-3088-4b38-8043-385246de2f06 · outbound

This paper cites Cognitive disentanglement for referring multi-object tracking.Information Fusion, page 103349.

ViewSAM: Learning View-aware Cross-modal Semantics for Weakly Supervised Cross-view Referring Multi-Object Tracking Cognitive disentanglement for referring multi-object tracking.Information Fusion, page 103349

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-05-26T04:46:43.285371Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-08T18:32:38.644948Z digest=sha256:7e90e5ee894f04038617c937fc5f2852052e2fad5207b63d7b76d53365d4147a

Observation 1c423cb9-723b-4e9e-8b3a-fe712fbab442 · outbound

This paper cites Visual-linguistic feature alignment with semantic and kinematic guidance for referring multi-object tracking.IEEE Transactions on Multimedia.

ViewSAM: Learning View-aware Cross-modal Semantics for Weakly Supervised Cross-view Referring Multi-Object Tracking Visual-linguistic feature alignment with semantic and kinematic guidance for referring multi-object tracking.IEEE Transactions on Multimedia

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-05-26T04:46:43.312439Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-08T18:32:38.644948Z digest=sha256:44b4bd223acf59fb5fe7dddfd94b24f77aa6e9210443a094e2ed88f6a5113456

Observation 6fc81412-d66f-4b21-b63d-41c75667cc03 · outbound

This paper cites SAMURAI: Adapting Segment Anything Model for Zero-Shot Visual Tracking with Motion-Aware Memory.

ViewSAM: Learning View-aware Cross-modal Semantics for Weakly Supervised Cross-view Referring Multi-Object Tracking SAMURAI: Adapting Segment Anything Model for Zero-Shot Visual Tracking with Motion-Aware Memory

Reference 34

Resolution
verified exact
arxiv_id, observed 2026-05-09T06:20:42.565832Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-08T18:32:38.644948Z digest=sha256:5fbab9e607c0f77943f1f7ddcea8bb41a1370de8b757d6b2e6ba133ba84b36de

Observation 36e3c076-a61f-407b-92ae-7463d76d1444 · outbound

This paper cites Sam2long: Enhancing sam 2 for long video segmentation with a training-free memory tree.

ViewSAM: Learning View-aware Cross-modal Semantics for Weakly Supervised Cross-view Referring Multi-Object Tracking Sam2long: Enhancing sam 2 for long video segmentation with a training-free memory tree

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-05-26T04:46:43.278261Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-08T18:32:38.644948Z digest=sha256:52de89a80f31a52150622dbd5084e4d2fb146f5b110a1a21c16dc6fe38c57d7a

Observation 0de1625c-8f40-4e15-9de5-fb88c4561d65 · outbound

This paper cites Sam2mot: A novel paradigm of multi-object tracking by segmentation.

ViewSAM: Learning View-aware Cross-modal Semantics for Weakly Supervised Cross-view Referring Multi-Object Tracking Sam2mot: A novel paradigm of multi-object tracking by segmentation

Reference 36

Resolution
verified exact
arxiv_id, observed 2026-05-09T06:20:42.572008Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-08T18:32:38.644948Z digest=sha256:170e9e8385d39ffd0cdee39e4fd3a0e2269caf977e43f50c43067d99b8b01369

Observation c2007e6d-bafb-459d-b75a-dffafe916708 · outbound

This paper cites Omni-scale feature learning for person re-identification.

ViewSAM: Learning View-aware Cross-modal Semantics for Weakly Supervised Cross-view Referring Multi-Object Tracking Omni-scale feature learning for person re-identification

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-05-26T04:46:43.282046Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-08T18:32:38.644948Z digest=sha256:57d7015f4cdd2da96a4a76586e6411a30a88b5aee5747eb844b73689ab439864

Observation c720873b-72a4-49ae-99bc-5d2bac1a6c08 · outbound

This paper cites Qwen3-VL Technical Report.

ViewSAM: Learning View-aware Cross-modal Semantics for Weakly Supervised Cross-view Referring Multi-Object Tracking Qwen3-VL Technical Report

Reference 38

Resolution
verified exact
local_arxiv, observed 2026-05-09T06:20:42.579877Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-08T18:32:38.644948Z digest=sha256:253aaaf4d3bc44a3c7b64effffe66e0950c7de16377ac6510834fba6e89f4e7a

Observation 6f10f714-f616-4b82-811a-8f6d90681e30 · outbound

This paper cites Parameter-efficient transfer learning for nlp.

ViewSAM: Learning View-aware Cross-modal Semantics for Weakly Supervised Cross-view Referring Multi-Object Tracking Parameter-efficient transfer learning for nlp

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-05-26T04:46:43.273744Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-08T18:32:38.644948Z digest=sha256:653035f0377b3aa28d4d7d52f363abfa129eb6ad10ac94f5c9de300c9748bedf

Observation f1ee9e6d-eb1a-4285-93e3-22526fb66d10 · outbound

This paper cites Film: Visual reasoning with a general conditioning layer.

ViewSAM: Learning View-aware Cross-modal Semantics for Weakly Supervised Cross-view Referring Multi-Object Tracking Film: Visual reasoning with a general conditioning layer

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-05-26T04:46:43.372379Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-08T18:32:38.644948Z digest=sha256:66793f18944eb5b12e9a6aaae25308dfff63b94a31c82ac67677454b0b590f7c

Observation fbb328d1-55bd-4825-9c1b-e7d8b7c3686b · outbound

This paper cites RoBERTa: A Robustly Optimized BERT Pretraining Approach.

ViewSAM: Learning View-aware Cross-modal Semantics for Weakly Supervised Cross-view Referring Multi-Object Tracking RoBERTa: A Robustly Optimized BERT Pretraining Approach

Reference 41

Resolution
verified exact
local_arxiv, observed 2026-05-09T06:20:42.587014Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-08T18:32:38.644948Z digest=sha256:6aa9e4e71f4fc2553ef3ad39e0409572ab8f250533dfef5970eae57f2222aac5

Observation c1b830b6-0a6c-4094-b473-4d89c8814ee0 · outbound

This paper cites Towards unified text-based person retrieval: A large-scale multi-attribute and language search benchmark.

ViewSAM: Learning View-aware Cross-modal Semantics for Weakly Supervised Cross-view Referring Multi-Object Tracking Towards unified text-based person retrieval: A large-scale multi-attribute and language search benchmark

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-05-26T04:46:43.315993Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-08T18:32:38.644948Z digest=sha256:997f369d8e0e3ce524656a118a7ffa1842d2466da3a7265e819b312e1f61bb7a

Observation 9100381e-44ff-48a4-b838-b79da8e5ebec · outbound

This paper cites Urvos: Unified referring video object segmentation network with a large-scale benchmark.

ViewSAM: Learning View-aware Cross-modal Semantics for Weakly Supervised Cross-view Referring Multi-Object Tracking Urvos: Unified referring video object segmentation network with a large-scale benchmark

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-05-26T04:46:43.344424Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-08T18:32:38.644948Z digest=sha256:2aa91882d403c9412ef16d7c65d83285936481fd4325a61ab81f4c36c845c128

Pith citing papers

No inbound Pith citation observations are available.