Pith. sign in

Paper Citation Record · LEDGER

Adaptive Perception for Unified Visual Multi-modal Object Tracking

As of 13 August 2026, this Paper Citation Record lists 60 of 60 outbound references and 1 inbound Pith citation observation for arXiv:2502.06583.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2502.06583 v1

Coverage vector

measured 60 of 60 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-08T15:01:38.486514Z

measured 61 of 61 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-13T06:32:02.005865+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-06T15:20:07.965249Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-06T15:20:13.646436Z

Reference resolution

60 of 60 outbound references displayed

  • verified exact6
  • verified fuzzy44
  • unresolved10
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation ae2b2bb3-95cb-4da3-8082-80449fcb96cb · outbound

This paper cites Backbone is all your need: A simplified architecture for visual object tracking,.

Adaptive Perception for Unified Visual Multi-modal Object Tracking Backbone is all your need: A simplified architecture for visual object tracking,

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T15:01:39.195544Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-08T15:01:38.256848Z digest=sha256:ce60e5ded9522647cf7d50f23515d6dff6d09d30eb72014bdce0ecfc57473294

Observation 6d37118a-2ee4-416a-bb42-55bddf870ca2 · outbound

This paper cites Mixformer: End-to-end tracking with iterative mixed attention,.

Adaptive Perception for Unified Visual Multi-modal Object Tracking Mixformer: End-to-end tracking with iterative mixed attention,

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T15:01:39.186036Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-08T15:01:38.261223Z digest=sha256:8c6ba4719d0c362e7934b0453595adfc21a125624e17e789016ba0d48718f985

Observation 774a935a-5458-49a9-9a30-c8e726e2e8e1 · outbound

This paper cites Joint feature learning and relation modeling for tracking: A one-stream framework,.

Adaptive Perception for Unified Visual Multi-modal Object Tracking Joint feature learning and relation modeling for tracking: A one-stream framework,

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T15:01:39.176512Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-08T15:01:38.264760Z digest=sha256:1bfb6f9b91ddec8b47097e3ca42c2c2697c69af8e5dce32a5c8cef3d950bb4a8

Observation c616e41c-04eb-485c-b31e-aaf76c4797ad · outbound

This paper cites Seqtrack: Sequence to sequence learning for visual object tracking,.

Adaptive Perception for Unified Visual Multi-modal Object Tracking Seqtrack: Sequence to sequence learning for visual object tracking,

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T15:01:39.166811Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-08T15:01:38.268681Z digest=sha256:395c090a48345c8d131e0173b63ed691abd9bdb1f3833493635fbc4a9ebc032e

Observation c4057e42-c032-4133-9eea-38776be93d97 · outbound

This paper cites Swintrack: A simple and strong baseline for transformer tracking,.

Adaptive Perception for Unified Visual Multi-modal Object Tracking Swintrack: A simple and strong baseline for transformer tracking,

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T15:01:39.156677Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-08T15:01:38.272577Z digest=sha256:4a6413821295d2dedba5178ed31f4dff3a545127d0ca8ed07c6fef828f1d3043

Observation f78f198c-5177-40cb-8af0-a877820e7249 · outbound

This paper cites Autoregressive visual tracking,.

Adaptive Perception for Unified Visual Multi-modal Object Tracking Autoregressive visual tracking,

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T15:01:39.146390Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-08T15:01:38.276091Z digest=sha256:c38b5b071281f7df8c038c23cbd18f96dd90d51b97c2858912fbf5d05ff6ccc8

Observation 8108e1d3-9c7a-4079-b285-da49a945e1df · outbound

This paper cites Visual prompt multi- modal tracking,.

Adaptive Perception for Unified Visual Multi-modal Object Tracking Visual prompt multi- modal tracking,

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T15:01:39.136277Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-08T15:01:38.280162Z digest=sha256:889081123e8c6bd509079df4341166363fe22355a6d08afadb9489d747f44c79

Observation f1e2e9a1-1af7-4cef-ac7c-22d2c45659a9 · outbound

This paper cites Single-Model and Any-Modality for Video Object Tracking.

Adaptive Perception for Unified Visual Multi-modal Object Tracking Single-Model and Any-Modality for Video Object Tracking

Reference 8

Resolution
verified exact
local_arxiv, observed 2026-08-08T15:01:38.664936Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-08T15:01:38.284479Z digest=sha256:f35420b527f97f11ca09307675fa0282da3aa710ae0da736243d01cb578c6b18

Observation e60ec162-1789-4f98-bd54-6dded71c3a08 · outbound

This paper cites Robust Tracking via Mamba-based Context-aware Token Learning.

Adaptive Perception for Unified Visual Multi-modal Object Tracking Robust Tracking via Mamba-based Context-aware Token Learning

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-08T15:01:38.289301Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T15:01:38.289301Z digest=sha256:3e6e92cde848d302294bdf7a545a16b6b6127510e92bbf719f284e3a2f179425

Observation dd84f229-f4bd-42ef-aead-49efbf9852fd · outbound

This paper cites Exploiting Multimodal Spatial-temporal Patterns for Video Object Tracking.

Adaptive Perception for Unified Visual Multi-modal Object Tracking Exploiting Multimodal Spatial-temporal Patterns for Video Object Tracking

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-08T15:01:38.293522Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T15:01:38.293522Z digest=sha256:10f69aa7252c2e28fa21ad7d9dcd36ffe4b9d4e7c21978240ad9b0780da532f7

Observation 58b423eb-40aa-447b-b708-d7faf8ac53cf · outbound

This paper cites Curricular contrastive regularization for physics-aware single image dehazing,.

Adaptive Perception for Unified Visual Multi-modal Object Tracking Curricular contrastive regularization for physics-aware single image dehazing,

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T15:01:39.125360Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-08T15:01:38.297829Z digest=sha256:72b4a0b09726524413df93928a605554973b75501c1b55dd94c4e66cf8b3293d

Observation 3127b424-4be4-4df1-a08f-b3163a397d14 · outbound

This paper cites Dynamic group difference coding based on thermal infrared face image for fever screening,.

Adaptive Perception for Unified Visual Multi-modal Object Tracking Dynamic group difference coding based on thermal infrared face image for fever screening,

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T15:01:39.114649Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-08T15:01:38.301951Z digest=sha256:311abb588e49c2cb99d2e28de57d646d646fb1cd8875333223765f425682eb44

Observation 3abafafd-98ca-473b-8504-fccd1ef64521 · outbound

This paper cites Few-shot learning with long- tailed labels,.

Adaptive Perception for Unified Visual Multi-modal Object Tracking Few-shot learning with long- tailed labels,

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T15:01:39.104340Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-08T15:01:38.305685Z digest=sha256:21a84fb7af2d13b7bfbc3d35c317784010aa3a049742788a9bf921595c0dc00f

Observation a59b6c1b-ce2c-4e4d-8279-c96f6390d2eb · outbound

This paper cites SHaRPose: Sparse High-Resolution Representation for Human Pose Estimation,.

Adaptive Perception for Unified Visual Multi-modal Object Tracking SHaRPose: Sparse High-Resolution Representation for Human Pose Estimation,

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T15:01:39.093728Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-08T15:01:38.309654Z digest=sha256:2f4f415b79f4676a0b768e906d36d3cc8f4161f8e18e5c19e16d0460e2546675

Observation ec604d4e-9319-411b-95c1-ff689e4564bf · outbound

This paper cites An emotion recognition method based on eye movement and audiovisual features in mooc learning environment,.

Adaptive Perception for Unified Visual Multi-modal Object Tracking An emotion recognition method based on eye movement and audiovisual features in mooc learning environment,

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T15:01:39.083502Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-08T15:01:38.313467Z digest=sha256:039ddfffda832c3b7984e2c330b7a3fc12c3a03651a33c93d8f6bae4f453fada

Observation 8c5c0319-ada0-4e34-926c-275af4dc4392 · outbound

This paper cites 3d- guided multi-feature semantic enhancement network for person re-id,.

Adaptive Perception for Unified Visual Multi-modal Object Tracking 3d- guided multi-feature semantic enhancement network for person re-id,

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T15:01:39.072896Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-08T15:01:38.317387Z digest=sha256:47ab1d3bbe36b0b43085df0f5b916e8b0c91e4cedf661d951e7ff64c42403d2d

Observation 8bc50acc-e8fb-43fc-99a0-b31dba64f507 · outbound

This paper cites Multi-branch enhanced discriminative network for vehicle re-identification,.

Adaptive Perception for Unified Visual Multi-modal Object Tracking Multi-branch enhanced discriminative network for vehicle re-identification,

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T15:01:39.061548Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-08T15:01:38.321291Z digest=sha256:40b04857d03f988930e13e0c5f1d3ede50316400a1d030f5d30c97d174374000

Observation 75e1fd3c-61ea-4eff-8542-49ef97f0067d · outbound

This paper cites Guided Real Image Dehazing using YCbCr Color Space.

Adaptive Perception for Unified Visual Multi-modal Object Tracking Guided Real Image Dehazing using YCbCr Color Space

Reference 18

Resolution
verified exact
local_arxiv, observed 2026-08-08T15:01:38.626079Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-08T15:01:38.325098Z digest=sha256:62fb850b16ed4084ebd7fb429e56a2c4e1eec3cc506d115528be87e67d4c4357

Observation b0a3ff3b-19c4-438f-97ed-c9fe611ac37c · outbound

This paper cites OneTracker: Unifying Visual Object Tracking with Foundation Models and Efficient Tuning.

Adaptive Perception for Unified Visual Multi-modal Object Tracking OneTracker: Unifying Visual Object Tracking with Foundation Models and Efficient Tuning

Reference 19

Resolution
verified exact
local_arxiv, observed 2026-08-08T15:01:38.612183Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-08T15:01:38.329265Z digest=sha256:d3a3311a65888dbb905e5206dc96b6e7aa2e3466655b4ca8f955d776c77d38ad

Observation b4afefca-df2a-48da-983e-0f1579b8a5fa · outbound

This paper cites Prompting for multi-modal tracking,.

Adaptive Perception for Unified Visual Multi-modal Object Tracking Prompting for multi-modal tracking,

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T15:01:39.050806Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-08T15:01:38.333793Z digest=sha256:febdf20db06e03eaefd9a461ff6f46d3ff5c42875f60645b98d482b2e3712c00

Observation 586c5bbc-acf0-469d-a51b-dce848553d4e · outbound

This paper cites SDSTrack: Self-Distillation Symmetric Adapter Learning for Multi-Modal Visual Object Tracking.

Adaptive Perception for Unified Visual Multi-modal Object Tracking SDSTrack: Self-Distillation Symmetric Adapter Learning for Multi-Modal Visual Object Tracking

Reference 21

Resolution
verified exact
local_arxiv, observed 2026-08-08T15:01:38.596444Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-08T15:01:38.338768Z digest=sha256:54706ee4b1a73bb206be693d3767bf0cc8dc65a5b50512990c467af8d233e3d9

Observation 2bc63b7a-a1ee-4a72-9de5-d6e6fbe98bda · outbound

This paper cites Bridging search region interaction with template for RGB-T tracking,.

Adaptive Perception for Unified Visual Multi-modal Object Tracking Bridging search region interaction with template for RGB-T tracking,

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T15:01:39.040569Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-08T15:01:38.343304Z digest=sha256:31105e0b4eb58ac9c1d2ab3ce6af797cb122d8338bb80afa63e632602f85a4f9

Observation 796e2777-6c9d-4570-bfd0-15090e63e4a1 · outbound

This paper cites Bi-directional adapter for multi- modal tracking,.

Adaptive Perception for Unified Visual Multi-modal Object Tracking Bi-directional adapter for multi- modal tracking,

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T15:01:39.029978Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-08T15:01:38.347554Z digest=sha256:581a1a44a4d24ea2b25f66e6d8804593165d7df09bdbd4937fddfc1a557e58fc

Observation 3ea5e07d-5d7e-4700-92e8-a06ce244ce02 · outbound

This paper cites Spiking transformers for event-based single object tracking,.

Adaptive Perception for Unified Visual Multi-modal Object Tracking Spiking transformers for event-based single object tracking,

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T15:01:39.020710Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-08T15:01:38.351534Z digest=sha256:c273836dd7095d3bfe6991dc005ea5ffe611f2a801a6b755cb7e24b8e8b2c3f6

Observation 8b1f559c-8ee2-4d90-bea0-50eae7c83025 · outbound

This paper cites Lasot: A high-quality benchmark for large-scale single object tracking,.

Adaptive Perception for Unified Visual Multi-modal Object Tracking Lasot: A high-quality benchmark for large-scale single object tracking,

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T15:01:39.009603Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-08T15:01:38.355559Z digest=sha256:c9cc6ab78a74025503785899ac9c78f793e4419d7fd1ece519ca6ded52f1b553

Observation c3d2d9ec-e421-4119-b0ba-26dbeb35d237 · outbound

This paper cites Got-10k: A large high-diversity benchmark for generic object tracking in the wild,.

Adaptive Perception for Unified Visual Multi-modal Object Tracking Got-10k: A large high-diversity benchmark for generic object tracking in the wild,

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T15:01:38.998252Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-08T15:01:38.359442Z digest=sha256:3f1adda44fc2a4bcbe235f004f6e7f49c25b0f719e62e757c4375947782d7565

Observation 194a8d21-c048-4cd1-8f62-cf8c9c9d92f2 · outbound

This paper cites Trackingnet: A large-scale dataset and benchmark for object tracking in the wild,.

Adaptive Perception for Unified Visual Multi-modal Object Tracking Trackingnet: A large-scale dataset and benchmark for object tracking in the wild,

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T15:01:38.986870Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-08T15:01:38.363352Z digest=sha256:40531dee1b683ab80ad8bbb0caef6dd9efde083a740b185456267bd1ee0bef82

Observation a73088f9-0cc1-41ce-8626-4b00a8f9b49c · outbound

This paper cites Lasher: A large-scale high-diversity benchmark for RGBT tracking,.

Adaptive Perception for Unified Visual Multi-modal Object Tracking Lasher: A large-scale high-diversity benchmark for RGBT tracking,

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T15:01:38.973857Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-08T15:01:38.367242Z digest=sha256:6548a9ccb8fe71c17fbe3547eb0297e8576dbe3742e76e6ef4c1c40e4081d070

Observation 96f5ea52-3207-43f2-8308-3c3de07a75ce · outbound

This paper cites RGB-T object tracking: Benchmark and baseline,.

Adaptive Perception for Unified Visual Multi-modal Object Tracking RGB-T object tracking: Benchmark and baseline,

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T15:01:38.961719Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-08T15:01:38.371104Z digest=sha256:e453c638b51b642b2ec2bbb224f09b34466b17fc5ead51fccc73049865ccae0f

Observation ab07304f-7740-4ad3-ab8b-7ecb33720098 · outbound

This paper cites Depthtrack: Unveiling the power of rgbd tracking,.

Adaptive Perception for Unified Visual Multi-modal Object Tracking Depthtrack: Unveiling the power of rgbd tracking,

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T15:01:38.949221Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-08T15:01:38.374858Z digest=sha256:6582693a70552e634eee67dd0c3a4c57cbc2d71d7c0665bb4ebc00bbd4983af7

Observation 3aa9dbad-8a55-40b9-93fd-6f09d8839568 · outbound

This paper cites The visual object tracking vot2015 challenge results,.

Adaptive Perception for Unified Visual Multi-modal Object Tracking The visual object tracking vot2015 challenge results,

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T15:01:38.936593Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-08T15:01:38.378540Z digest=sha256:503af260b0c22e49da440604a80857009f777b768e888a6185a47aa14a7482de

Observation 88f7be3a-626a-4d16-8af3-b4b9962248e6 · outbound

This paper cites Visevent: Reliable object tracking via collaboration of frame and event flows,.

Adaptive Perception for Unified Visual Multi-modal Object Tracking Visevent: Reliable object tracking via collaboration of frame and event flows,

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T15:01:38.924562Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-08T15:01:38.382631Z digest=sha256:be7b3652d8c167fb91aed8a5a386058c91dd4f3551dc0b1b585b7f6a6020d823

Observation 6115b44c-0ead-413f-9f68-3ca71ac655d3 · outbound

This paper cites Unified-io: A unified model for vision, language, and multi-modal tasks,.

Adaptive Perception for Unified Visual Multi-modal Object Tracking Unified-io: A unified model for vision, language, and multi-modal tasks,

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-08T15:01:38.386671Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T15:01:38.386671Z digest=sha256:8bf47be07fa72c70c93566edd7107905d6e2d7f3fb3bfda0854a37965422ba2c

Observation 59d77813-4c96-4129-8b07-9a6faae8009e · outbound

This paper cites Imagebind: One embedding space to bind them all,.

Adaptive Perception for Unified Visual Multi-modal Object Tracking Imagebind: One embedding space to bind them all,

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T15:01:38.904614Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-08T15:01:38.390455Z digest=sha256:d01e9dfaf75045a49fdaed8182bac6b8825277f646f25f76cf4b8e28076f63f6

Observation a63fec7a-da80-4305-abd7-d572bb5a9d0d · outbound

This paper cites MUTEX: Learning Unified Policies from Multimodal Task Specifications.

Adaptive Perception for Unified Visual Multi-modal Object Tracking MUTEX: Learning Unified Policies from Multimodal Task Specifications

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-08T15:01:38.393426Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T15:01:38.393426Z digest=sha256:2434749a1028c3c25812af227f794c7f60b0669bc4d156afebe7c4d6d418373c

Observation 766d9f0a-ffd3-4334-8fdf-9469ea52f920 · outbound

This paper cites Siamese Vision Transformers are Scalable Audio-visual Learners.

Adaptive Perception for Unified Visual Multi-modal Object Tracking Siamese Vision Transformers are Scalable Audio-visual Learners

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-08T15:01:38.397539Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T15:01:38.397539Z digest=sha256:0c591c9470109f5bb8a0bb692765164145c184944202170fdf02bd76be0f1c2c

Observation f2bfc63a-4f65-4813-bcf2-2e19670f1d95 · outbound

This paper cites A unified audio-visual learning framework for localization, separation, and recognition,.

Adaptive Perception for Unified Visual Multi-modal Object Tracking A unified audio-visual learning framework for localization, separation, and recognition,

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T15:01:38.891981Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-08T15:01:38.401184Z digest=sha256:68f6b999ebeae0b3615692f5779d93318b8eee69f9ff6616855bd670751ffd82

Observation 0c1a4fa6-a92c-4773-85e1-692bb4cdda91 · outbound

This paper cites Learning visual representation from modality-shared contrastive language-image pre-training,.

Adaptive Perception for Unified Visual Multi-modal Object Tracking Learning visual representation from modality-shared contrastive language-image pre-training,

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T15:01:38.878140Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-08T15:01:38.404589Z digest=sha256:169a64ef773bba60e4b36be7555c1d4315a0e02b2f244c4ae556e0521da5367e

Observation cdfd3f0b-d730-4075-b0d0-e834ec74cfb6 · outbound

This paper cites Decoupled weight decay regularization,.

Adaptive Perception for Unified Visual Multi-modal Object Tracking Decoupled weight decay regularization,

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T15:01:38.866174Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-08T15:01:38.407603Z digest=sha256:346dad6bad94c001a9ab52a04fea6fa7e9177c72268bc6a722ead93bc8c49251

Observation 5942d894-1489-408f-988d-916e45a4d3db · outbound

This paper cites Weighted sparse representa- tion regularized graph learning for rgb-t object tracking,.

Adaptive Perception for Unified Visual Multi-modal Object Tracking Weighted sparse representa- tion regularized graph learning for rgb-t object tracking,

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T15:01:38.855104Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-08T15:01:38.410684Z digest=sha256:ecd8cb3c8f8a170f69ba82e19cd80d1fb2c32c418e273c07bffcab448f9d9794

Observation 0371c9fe-b5e7-4fc6-856f-357ab1008aa2 · outbound

This paper cites Generative-based fusion mechanism for multi-modal tracking,.

Adaptive Perception for Unified Visual Multi-modal Object Tracking Generative-based fusion mechanism for multi-modal tracking,

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T15:01:38.845254Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-08T15:01:38.413706Z digest=sha256:c444980bd19b3aabc8962ca5b76a46e47017ce141b507613f8095fef498d721e

Observation 59dbba3c-c935-4ac3-b145-53800b2d829f · outbound

This paper cites Transformer tracking,.

Adaptive Perception for Unified Visual Multi-modal Object Tracking Transformer tracking,

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T15:01:38.835252Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-08T15:01:38.416861Z digest=sha256:5d4d08725a94b4806a876b84f509657f607c7d3062519e4a7a8e37692796bd1b

Observation e0c0914f-3e9b-45e1-8846-ba04d4acff3b · outbound

This paper cites Learning spatio-temporal transformer for visual tracking,.

Adaptive Perception for Unified Visual Multi-modal Object Tracking Learning spatio-temporal transformer for visual tracking,

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-08T15:01:38.420525Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T15:01:38.420525Z digest=sha256:598cde20fc75d83c68d185c98b38551063d4d48c76128890f19b3af343aa9318

Observation e4ff1443-2a31-4fbf-be4a-4dacc6dfc7e1 · outbound

This paper cites Aiatrack: Attention in attention for transformer visual tracking,.

Adaptive Perception for Unified Visual Multi-modal Object Tracking Aiatrack: Attention in attention for transformer visual tracking,

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-08T15:01:38.424619Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T15:01:38.424619Z digest=sha256:c4e3b7439ee1287cd54a0b1d769b5a3535a5d131cf36dbd4ad1316e1694d150b

Observation 8e89baee-71c9-431c-8944-3a5acd28a9be · outbound

This paper cites Rgbd1k: A large-scale dataset and benchmark for rgb-d object tracking,.

Adaptive Perception for Unified Visual Multi-modal Object Tracking Rgbd1k: A large-scale dataset and benchmark for rgb-d object tracking,

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T15:01:38.812467Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-08T15:01:38.428469Z digest=sha256:ef23d943cc52a3dfacc4a182983b0fdbc6e8f59a493c86b6f426a4f89a9fe9d8

Observation b0677cfa-d238-45bc-aacf-7fea2b8c6926 · outbound

This paper cites Focal loss for dense object detection,.

Adaptive Perception for Unified Visual Multi-modal Object Tracking Focal loss for dense object detection,

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-08T15:01:38.432244Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T15:01:38.432244Z digest=sha256:a02cd0dfbf886bc65bbd3d992f2ed52b027178f94bfeb15fabdb6e9a237f618f

Observation 823456f3-4507-4856-9a7e-96e38f8ab8fb · outbound

This paper cites Generalized intersection over union: A metric and a loss for bounding box regression,.

Adaptive Perception for Unified Visual Multi-modal Object Tracking Generalized intersection over union: A metric and a loss for bounding box regression,

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T15:01:38.794074Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-08T15:01:38.436214Z digest=sha256:77996823cb2f0fe0c3ab9284ebde192589ebff664814ed55cc9edff96f9a432f

Observation daa29079-8457-4bd8-9f6c-50806d158431 · outbound

This paper cites Depthtrack: Unveiling the power of RGBD tracking,.

Adaptive Perception for Unified Visual Multi-modal Object Tracking Depthtrack: Unveiling the power of RGBD tracking,

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T15:01:38.782160Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-08T15:01:38.440102Z digest=sha256:bf4b9bb31f3fc285b20e03747c4855b15612bc1e75ad234c00f8416ad1efc932

Observation d5690a09-d456-4f20-ab09-aecd738e4914 · outbound

This paper cites Transformer tracking via frequency fusion,.

Adaptive Perception for Unified Visual Multi-modal Object Tracking Transformer tracking via frequency fusion,

Reference 49

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T15:01:38.770507Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-08T15:01:38.443660Z digest=sha256:fab5153833561c08d4caf4241ec0c5af86d069f32b5265781c4853e5517f27a1

Observation 33dce51f-e2f7-4125-a47a-ad004815f6a1 · outbound

This paper cites Multiple source domain adaptation for multiple object tracking in satellite video,.

Adaptive Perception for Unified Visual Multi-modal Object Tracking Multiple source domain adaptation for multiple object tracking in satellite video,

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T15:01:38.758488Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-08T15:01:38.447390Z digest=sha256:879867ba111d850dc7f98110ee18705e59c91ed961051d7c56dccb4c7f77ec00

Observation 7dd02a68-a0b1-451a-9eea-7850cf970597 · outbound

This paper cites Explicit visual prompts for visual object tracking,.

Adaptive Perception for Unified Visual Multi-modal Object Tracking Explicit visual prompts for visual object tracking,

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T15:01:38.747050Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-08T15:01:38.451209Z digest=sha256:aee8ebc797fdfe4d000745b31125d06bee2203a2a74df2f285868b26da9a1760

Observation 6b9fe4db-d580-487f-96d2-73ff40e0dd65 · outbound

This paper cites Autoregressive queries for adaptive tracking with spatio-temporal trans- formers,.

Adaptive Perception for Unified Visual Multi-modal Object Tracking Autoregressive queries for adaptive tracking with spatio-temporal trans- formers,

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-08T15:01:38.455226Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T15:01:38.455226Z digest=sha256:4c05ee5f254b5381b99c5ac41d7130b85fe03775c59359d164d7aab47c47a07e

Observation fcdd4def-e275-4da9-9e85-5a9eeb715089 · outbound

This paper cites Learning transferable visual models from natural language supervision,.

Adaptive Perception for Unified Visual Multi-modal Object Tracking Learning transferable visual models from natural language supervision,

Reference 53

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T15:01:38.726920Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-08T15:01:38.458954Z digest=sha256:10cc9a1d623e002db9e185e92e4b1055865e7f0d6b2785a23cd2539b1dda0427

Observation 83c21fd7-bdef-41bb-8554-d16961611ed2 · outbound

This paper cites Towards modalities correlation for rgb-t tracking,.

Adaptive Perception for Unified Visual Multi-modal Object Tracking Towards modalities correlation for rgb-t tracking,

Reference 54

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T15:01:38.713885Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-08T15:01:38.462662Z digest=sha256:1e6bbfee8b47d18b902f27f94d08c4e57ca6f31d77146b1543e6a808268c94e8

Observation d9553625-d158-4c93-a018-9920580c311d · outbound

This paper cites RGBD1K: A large-scale dataset and benchmark for RGB-D object tracking,.

Adaptive Perception for Unified Visual Multi-modal Object Tracking RGBD1K: A large-scale dataset and benchmark for RGB-D object tracking,

Reference 55

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T15:01:38.702132Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-08T15:01:38.466373Z digest=sha256:e34f687761ebbabd7eb1d195353a3bb8583397fabcf2029857cc609995acaa50

Observation 64227ba3-8b07-4445-96dc-68280b42ce2f · outbound

This paper cites Siamban: Target-aware tracking with siamese box adaptive network,.

Adaptive Perception for Unified Visual Multi-modal Object Tracking Siamban: Target-aware tracking with siamese box adaptive network,

Reference 56

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T15:01:38.689883Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-08T15:01:38.470170Z digest=sha256:01ba978c13bfc436f014a66a867387918c1201b69744f80130238bfc4ce53f1b

Observation f0f4888b-998b-4d60-ac2b-34faa6667c2a · outbound

This paper cites Reliable Object Tracking by Multimodal Hybrid Feature Extraction and Transformer-Based Fusion.

Adaptive Perception for Unified Visual Multi-modal Object Tracking Reliable Object Tracking by Multimodal Hybrid Feature Extraction and Transformer-Based Fusion

Reference 57

Resolution
verified exact
local_arxiv, observed 2026-08-08T15:01:38.557849Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-08T15:01:38.473979Z digest=sha256:841572239f8507a6de23251d3393069c10923097842c17630c018184a6d6fb24

Observation 63703f02-2ee1-466c-a648-9804ce6f755a · outbound

This paper cites Depthrefiner: Adapting rgb trackers to rgbd scenes via depth-fused refinement,.

Adaptive Perception for Unified Visual Multi-modal Object Tracking Depthrefiner: Adapting rgb trackers to rgbd scenes via depth-fused refinement,

Reference 58

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T15:01:38.677307Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-08T15:01:38.478408Z digest=sha256:05c12fc01e72104d0052ebe13b1218e809bedf73e894aabbf44a4394cf10b405

Observation 1ab43ecb-3c4d-4d1e-b0cb-799a3e344e0e · outbound

This paper cites Cross-modulated Attention Transformer for RGBT Tracking.

Adaptive Perception for Unified Visual Multi-modal Object Tracking Cross-modulated Attention Transformer for RGBT Tracking

Reference 59

Resolution
verified exact
local_arxiv, observed 2026-08-08T15:01:38.539120Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-08T15:01:38.482277Z digest=sha256:fa7aa852bfed5a721d3f817e58fa393470b403e47651475d52d0fa99ccadbaf4

Observation 876085fb-394e-4348-8dda-c6647c7db7cc · outbound

This paper cites RGBT Tracking via All-layer Multimodal Interactions with Progressive Fusion Mamba.

Adaptive Perception for Unified Visual Multi-modal Object Tracking RGBT Tracking via All-layer Multimodal Interactions with Progressive Fusion Mamba

Reference 60

Resolution
unresolved
no resolver link, observed 2026-08-08T15:01:38.486514Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T15:01:38.486514Z digest=sha256:9c88933f3987487dc9767f526573eb3921cb1256c4950c0510640781fb6609ed

Pith citing papers

Observation 16c70ea3-faff-4ac1-8bd3-9e1cdca7e96a · inbound

Explicit Context Reasoning with Supervision for Visual Tracking cites this paper.

Explicit Context Reasoning with Supervision for Visual Tracking Adaptive Perception for Unified Visual Multi-modal Object Tracking

Reference 18

Resolution
verified exact
local_arxiv, observed 2026-08-06T15:20:13.728895Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-06T15:20:07.965249Z digest=sha256:337fe5f2f044e0eb223f12dcd3583d3fad1b963ab3f2b14dd9e1ec31737ca37b