Pith. sign in

Paper Citation Record · LEDGER

UNIV: Unified Foundation Model for Infrared and Visible Modalities

As of 19 August 2026, this Paper Citation Record lists 45 of 45 outbound references and 1 inbound Pith citation observation for arXiv:2509.15642.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2509.15642 v3

Coverage vector

measured 45 of 45 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-05-18T16:28:48.494939Z

measured 46 of 46 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-18T06:34:40.430872+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-01T10:27:20.578105Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

45 of 45 outbound references displayed

  • verified exact9
  • verified fuzzy35
  • unresolved1
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 7b883189-b104-42b9-a5d1-35a47287ce9f · outbound

This paper cites https://github.com/ultralytics/ultralytics.

UNIV: Unified Foundation Model for Infrared and Visible Modalities https://github.com/ultralytics/ultralytics

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T16:31:38.564399Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-18T16:28:48.494939Z digest=sha256:0fc29f9bf942bc49b699f2d6fc8698e054fa105f8cca5e33e989f861abdb0225

Observation 80dd73f5-dbb2-454c-a8e1-92152eaf5756 · outbound

This paper cites Self-supervised learning from images with a joint-embedding predictive architecture.

UNIV: Unified Foundation Model for Infrared and Visible Modalities Self-supervised learning from images with a joint-embedding predictive architecture

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T16:31:38.576943Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-18T16:28:48.494939Z digest=sha256:a2d35dd91f6afef6650a5d3cd77479c6d3dfab3457c38da4df414b01449a6463

Observation fc574d61-9aea-4975-b4f1-62b4c9830e81 · outbound

This paper cites End- to-end object detection with transformers.

UNIV: Unified Foundation Model for Infrared and Visible Modalities End- to-end object detection with transformers

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T16:31:38.579698Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-18T16:28:48.494939Z digest=sha256:f7b0f8756f4e8c31e8ebbc82034945ca3b95402e35fddda39561f103363e83e6

Observation a526a00f-a6b0-46a8-ac96-10b9b0767899 · outbound

This paper cites Emerg- ing properties in self-supervised vision transformers.

UNIV: Unified Foundation Model for Infrared and Visible Modalities Emerg- ing properties in self-supervised vision transformers

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T16:31:38.582665Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-18T16:28:48.494939Z digest=sha256:a8a23da376478ee7ce0b402f95d90328d3320e60aeaeba17e601ad6326b5b431

Observation ff014e28-f218-460a-a748-ec74bab49fcc · outbound

This paper cites Encoder-decoder with atrous separable convolution for semantic image segmentation.

UNIV: Unified Foundation Model for Infrared and Visible Modalities Encoder-decoder with atrous separable convolution for semantic image segmentation

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T16:31:38.573402Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-18T16:28:48.494939Z digest=sha256:f5b08e27321c6a8a61dc7b057ae914926d21a78d4123d8558633ac3a88ee9182

Observation 7302c942-5950-499b-a37c-041edcc9d880 · outbound

This paper cites An empiri- cal study of training self-supervised vision transformers.

UNIV: Unified Foundation Model for Infrared and Visible Modalities An empiri- cal study of training self-supervised vision transformers

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T16:31:38.561437Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-18T16:28:48.494939Z digest=sha256:7b43135260f99a23e4ccf548f3b8f308fce87f1a3c94fc2b83700755fb9ed382

Observation 5ba99090-640a-4adc-9bab-5f7657fcc4db · outbound

This paper cites Tagging before alignment: Integrating multi- modal tags for video-text retrieval.

UNIV: Unified Foundation Model for Infrared and Visible Modalities Tagging before alignment: Integrating multi- modal tags for video-text retrieval

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T16:31:38.570775Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-18T16:28:48.494939Z digest=sha256:2943854561351550aedc747a3c8e1a56c71dd7bcba29f346bd59b260630639c6

Observation c33a4ad2-9401-4d77-935b-1f9285ffaea0 · outbound

This paper cites Learning a similarity metric discriminatively, with application to face verification.

UNIV: Unified Foundation Model for Infrared and Visible Modalities Learning a similarity metric discriminatively, with application to face verification

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T16:31:38.567439Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-18T16:28:48.494939Z digest=sha256:a4f6ddffd597594c8d6fee7e37129984196707bb5cf2fe45f33ef6e05dcb4f40

Observation e6dedf02-82dc-41ff-8e79-68d447217abd · outbound

This paper cites Imagenet: A large-scale hierarchical image database.

UNIV: Unified Foundation Model for Infrared and Visible Modalities Imagenet: A large-scale hierarchical image database

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T16:31:38.558694Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-18T16:28:48.494939Z digest=sha256:ec4bba31e29031dc76c531629d369514c8528046f3634aa4592f39d48ce6c948

Observation 4ff63ccc-9812-40a3-bf1f-492d26fa70d9 · outbound

This paper cites BERT: Pre-training of Deep Bidirectional Transformers for Language Understanding.

UNIV: Unified Foundation Model for Infrared and Visible Modalities BERT: Pre-training of Deep Bidirectional Transformers for Language Understanding

Reference 10

Resolution
verified exact
local_arxiv, observed 2026-05-18T16:31:37.143616Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-18T16:28:48.494939Z digest=sha256:05f47b1c69e69ed1ed8697ce09eb0a8af8410515f01405eb4e2857e5eae42f07

Observation 4b2b1fdb-cb07-44e4-ae85-d5fc04451865 · outbound

This paper cites An Image is Worth 16x16 Words: Transformers for Image Recognition at Scale.

UNIV: Unified Foundation Model for Infrared and Visible Modalities An Image is Worth 16x16 Words: Transformers for Image Recognition at Scale

Reference 11

Resolution
verified exact
local_arxiv, observed 2026-05-18T16:31:37.180208Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-18T16:28:48.494939Z digest=sha256:87c3e68b06d7c4fd3d4ad8c85b6735831dbf5feceae0640503e6e1e122b2bd06

Observation 95a9f9d2-43b0-43f6-8dd1-3f4dac6969a9 · outbound

This paper cites Flir thermal dataset version 1.3 [dataset].

UNIV: Unified Foundation Model for Infrared and Visible Modalities Flir thermal dataset version 1.3 [dataset]

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T16:32:45.373282Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-18T16:28:48.494939Z digest=sha256:f9ade828463c51aad05c481587d902d8879fda53d4c485dbf775c501e81e69d5

Observation fd48258d-c686-4772-aeea-480187373199 · outbound

This paper cites an unresolved cited work.

UNIV: Unified Foundation Model for Infrared and Visible Modalities Unresolved cited work

Reference 13

Resolution
unresolved
raw_fallback, observed 2026-05-18T16:32:45.391805Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-18T16:28:48.494939Z digest=sha256:ec6e5a7712a6c37734c6cd243fa2de6e0df1b178dd0c1479d2f1a4b0aab3db01

Observation 4fb88795-19fd-455e-a07d-0bbefa7858d2 · outbound

This paper cites Catastrophic forgetting in connectionist networks.Trends in cognitive sciences, 3(4):128–135.

UNIV: Unified Foundation Model for Infrared and Visible Modalities Catastrophic forgetting in connectionist networks.Trends in cognitive sciences, 3(4):128–135

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T16:32:45.384910Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-18T16:28:48.494939Z digest=sha256:063ed1b22717224d93e47bc921accd05f50c8f15884781580c3d362309f06091

Observation 53301122-e33a-4ca9-8e36-8d2968edc846 · outbound

This paper cites ConvMAE: Masked Convolution Meets Masked Autoencoders.

UNIV: Unified Foundation Model for Infrared and Visible Modalities ConvMAE: Masked Convolution Meets Masked Autoencoders

Reference 15

Resolution
verified exact
arxiv_id, observed 2026-05-18T16:31:37.170898Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-18T16:28:48.494939Z digest=sha256:e7cae191f38ce80c938986ebbd8829858010caf5c07b8a38c3557c8389516306

Observation 3b11516e-3565-4532-81b2-48b21306eb97 · outbound

This paper cites Mfnet: Towards real-time se- mantic segmentation for autonomous vehicles with multi- spectral scenes.

UNIV: Unified Foundation Model for Infrared and Visible Modalities Mfnet: Towards real-time se- mantic segmentation for autonomous vehicles with multi- spectral scenes

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T16:32:45.401635Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-18T16:28:48.494939Z digest=sha256:a153dd7d432ccdd9d6f1a62633b515e2bac8f575dbe63670d23aa616195c30b9

Observation 129afbed-c491-4414-b7ad-14c51a39ea75 · outbound

This paper cites Mask r-cnn.

UNIV: Unified Foundation Model for Infrared and Visible Modalities Mask r-cnn

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T16:32:45.398520Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-18T16:28:48.494939Z digest=sha256:15729971250a16c0d19b150e071e7a84d90c1ec278e15392b23192b226395013

Observation 499f363c-40fe-489a-a385-fa2c98a46f06 · outbound

This paper cites Masked autoencoders are scalable vision learners.

UNIV: Unified Foundation Model for Infrared and Visible Modalities Masked autoencoders are scalable vision learners

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T16:32:45.368227Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-18T16:28:48.494939Z digest=sha256:fa742bb9b55ab48f54f80415e0298153d6d695cc55956199d56a6312302180d2

Observation d694142c-0bbd-44cf-974d-9455684d00d3 · outbound

This paper cites Vic-mae: Self-supervised representation learning from im- ages and video with contrastive masked autoencoders.

UNIV: Unified Foundation Model for Infrared and Visible Modalities Vic-mae: Self-supervised representation learning from im- ages and video with contrastive masked autoencoders

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T16:32:45.351203Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-18T16:28:48.494939Z digest=sha256:d0e8f9c1ed414cad3c3b75efa468aba61b719311d63ee5c87c78f29e9fa12311

Observation 539bfa7c-0004-4261-a864-e53e8b78d681 · outbound

This paper cites MILAN: Masked Image Pretraining on Language Assisted Representation.

UNIV: Unified Foundation Model for Infrared and Visible Modalities MILAN: Masked Image Pretraining on Language Assisted Representation

Reference 20

Resolution
verified exact
arxiv_id, observed 2026-05-18T16:31:37.148483Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-18T16:28:48.494939Z digest=sha256:3cae99ecd5d7d576a75fa051151a8202091a8df24ed386d83c7461b02c09d4c7

Observation 2fc8c05f-8cba-498c-93cf-0c9663f1fd1d · outbound

This paper cites Interlaced Sparse Self-Attention for Semantic Segmentation.

UNIV: Unified Foundation Model for Infrared and Visible Modalities Interlaced Sparse Self-Attention for Semantic Segmentation

Reference 21

Resolution
verified exact
arxiv_id, observed 2026-05-18T16:31:37.161675Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-18T16:28:48.494939Z digest=sha256:44afa2d7f12ab03bbce3d724377e2a2aa26e7f9941fa36eab23b793b115d849e

Observation 670986c7-6e40-4952-979e-7b0a041f4b73 · outbound

This paper cites Multispectral pedestrian detection: Benchmark dataset and baseline.

UNIV: Unified Foundation Model for Infrared and Visible Modalities Multispectral pedestrian detection: Benchmark dataset and baseline

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T16:32:45.354290Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-18T16:28:48.494939Z digest=sha256:d0369485604f48c570fa64cdc9f701e60f05ac51ab66a7a4f13a727bfa538d69

Observation 0ae2a3c4-5690-49d9-b3b6-153260e0064f · outbound

This paper cites Llvip: A visible-infrared paired dataset for low-light vision.

UNIV: Unified Foundation Model for Infrared and Visible Modalities Llvip: A visible-infrared paired dataset for low-light vision

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T16:32:45.317229Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-18T16:28:48.494939Z digest=sha256:4b9a65d6261cae0e75f71972322a0b88a820fbacc6fde8c00e64878a3431a6f3

Observation a47151a0-7d07-4b6b-b746-a5aa31a36cd9 · outbound

This paper cites Tencent text-video retrieval: hierarchi- cal cross-modal interactions with multi-level representations.

UNIV: Unified Foundation Model for Infrared and Visible Modalities Tencent text-video retrieval: hierarchi- cal cross-modal interactions with multi-level representations

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T16:32:45.357759Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-18T16:28:48.494939Z digest=sha256:ed43cbff233ce87439fbde09e75e1250dc55a2d2e2094c8e31c1b7cfabd5ac43

Observation 9849a1d1-0c09-485e-b2e2-114b3c0f1c3a · outbound

This paper cites Segmenting objects in day and night: Edge-conditioned cnn for thermal image semantic segmentation.TNNLS, 32(7): 3069–3082.

UNIV: Unified Foundation Model for Infrared and Visible Modalities Segmenting objects in day and night: Edge-conditioned cnn for thermal image semantic segmentation.TNNLS, 32(7): 3069–3082

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T16:32:45.388556Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-18T16:28:48.494939Z digest=sha256:b451ced266349cb712d9e1a61450b4beb18fa6c9a03d14d07e198aa5119c4f92

Observation bba935d6-8cb0-4e9c-9254-918d57dc825e · outbound

This paper cites Lasher: A large-scale high- diversity benchmark for rgbt tracking.TIP, 31:392–404.

UNIV: Unified Foundation Model for Infrared and Visible Modalities Lasher: A large-scale high- diversity benchmark for rgbt tracking.TIP, 31:392–404

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T16:32:45.321893Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-18T16:28:48.494939Z digest=sha256:eb5452feac499002c8ec104640069d561335ce3d3d3ff0e5cd5e2991012937b4

Observation 977d941b-9127-4cca-9e4b-56d854c83d9d · outbound

This paper cites Feature pyramid networks for object detection.

UNIV: Unified Foundation Model for Infrared and Visible Modalities Feature pyramid networks for object detection

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T16:32:45.377500Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-18T16:28:48.494939Z digest=sha256:6e0a3f2adbd8d0931485fe7a45f4590f1437d605c93c9bc8d289987c35e927a3

Observation ecb22407-e09d-44e5-8a50-0ef0665779ea · outbound

This paper cites Infmae: A foundation model in the infrared modality.

UNIV: Unified Foundation Model for Infrared and Visible Modalities Infmae: A foundation model in the infrared modality

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T16:32:45.337725Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-18T16:28:48.494939Z digest=sha256:1ebd35739cd3d33b4fc5f3ddca269adb2257dfde5b62a9391ab279e291e6396a

Observation 9c8f12f3-e134-4288-9aab-f4e382cf4c66 · outbound

This paper cites Target-aware dual adversarial learning and a multi-scenario multi-modality benchmark to fuse infrared and visible for object detection.

UNIV: Unified Foundation Model for Infrared and Visible Modalities Target-aware dual adversarial learning and a multi-scenario multi-modality benchmark to fuse infrared and visible for object detection

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T16:32:45.345230Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-18T16:28:48.494939Z digest=sha256:cae0c459cff2c1fb672eb76b5529529059838f4cdf3089207e1196c8e5b0f6ab

Observation 40059d25-3d69-4b6c-8f97-7356b393403e · outbound

This paper cites X-clip: End-to-end multi-grained con- trastive learning for video-text retrieval.

UNIV: Unified Foundation Model for Infrared and Visible Modalities X-clip: End-to-end multi-grained con- trastive learning for video-text retrieval

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T16:32:45.330038Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-18T16:28:48.494939Z digest=sha256:1671de20f4cbdf5001f4d5a0fd2bff18a1c6196cd4f77740ccaae13718074e02

Observation c14689af-6625-4c21-9dd5-2a781b3a2e02 · outbound

This paper cites Multi- level cross-modal semantic alignment network for video–text retrieval.Mathematics, 10(18):3346.

UNIV: Unified Foundation Model for Infrared and Visible Modalities Multi- level cross-modal semantic alignment network for video–text retrieval.Mathematics, 10(18):3346

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T16:32:45.333648Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-18T16:28:48.494939Z digest=sha256:1f0c7e8835cfb3a059bd44b140b21d67d43f7136f939ce8a98cc4db6380216fa

Observation 9bd0514a-bca2-42ce-b79d-67fde6eebb96 · outbound

This paper cites DINOv2: Learning Robust Visual Features without Supervision.

UNIV: Unified Foundation Model for Infrared and Visible Modalities DINOv2: Learning Robust Visual Features without Supervision

Reference 32

Resolution
verified exact
local_arxiv, observed 2026-05-18T16:31:37.152875Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-18T16:28:48.494939Z digest=sha256:473aca9319803ab3fba92cf746cb577f040098b5fa834e1d7447d16106315c18

Observation 82c40e70-4f60-47a8-a1df-8dbb573fb9df · outbound

This paper cites Learn- ing transferable visual models from natural language super- vision.

UNIV: Unified Foundation Model for Infrared and Visible Modalities Learn- ing transferable visual models from natural language super- vision

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T16:31:38.588780Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-18T16:28:48.494939Z digest=sha256:247447546566955ab41e35584b3874afd1f87ae594a71425b8d0090fc1b725b1

Observation c225f6c9-7854-4ae2-9c34-395fdcff2f90 · outbound

This paper cites Faster r-cnn: Towards real-time object detection with region proposal networks.PAMI, 39(6):1137–1149.

UNIV: Unified Foundation Model for Infrared and Visible Modalities Faster r-cnn: Towards real-time object detection with region proposal networks.PAMI, 39(6):1137–1149

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T16:32:45.341435Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-18T16:28:48.494939Z digest=sha256:ffc44902e11bb16d6aa0f7840828d833dc40896c941c1faf73e96efac2fe8f8c

Observation 212aae39-2afd-4154-81aa-c65fd10cd165 · outbound

This paper cites The earth mover’s distance as a metric for image retrieval.In- ternational journal of computer vision, 40(2):99–121.

UNIV: Unified Foundation Model for Infrared and Visible Modalities The earth mover’s distance as a metric for image retrieval.In- ternational journal of computer vision, 40(2):99–121

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T16:31:38.592078Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-18T16:28:48.494939Z digest=sha256:cd1ab3fbd9e5b5598ab258963ebb01f019189ebfacd7a1a0b97d16a2aa7822d1

Observation 2bcf8726-9a65-4f41-b42a-7d74d21c2c20 · outbound

This paper cites Drone-based rgb-infrared cross-modality vehicle detection via uncertainty-aware learning.TCSVT, 32(10):6700–6713.

UNIV: Unified Foundation Model for Infrared and Visible Modalities Drone-based rgb-infrared cross-modality vehicle detection via uncertainty-aware learning.TCSVT, 32(10):6700–6713

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T16:31:38.585623Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-18T16:28:48.494939Z digest=sha256:2c56bab7e7ee88ce3f9f6b0edab6ac61b7f61b87312d7efe3dd82348da7b42dc

Observation 8f350059-768c-4a97-beb8-26d99a61f0e0 · outbound

This paper cites Piafusion: A progressive infrared and visible im- age fusion network based on illumination aware.Information Fusion.

UNIV: Unified Foundation Model for Infrared and Visible Modalities Piafusion: A progressive infrared and visible im- age fusion network based on illumination aware.Information Fusion

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T16:31:38.595180Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-18T16:28:48.494939Z digest=sha256:96363b082403d2e0c454c09138efc09b95576a69c17e093b8e538d50b8a935bc

Observation 60b6caac-a745-4695-b6ea-4b0dbbd09b16 · outbound

This paper cites Fine-grained action retrieval through multiple parts- of-speech embeddings.

UNIV: Unified Foundation Model for Infrared and Visible Modalities Fine-grained action retrieval through multiple parts- of-speech embeddings

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T16:32:45.326472Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-18T16:28:48.494939Z digest=sha256:9a187f7a5476c2b56d108cf0836ee3ab56cb8896bbec3bcd6be184269c464087

Observation 40e74b97-18ab-4f2e-b782-b815d7afca10 · outbound

This paper cites Unified perceptual parsing for scene understand- ing.

UNIV: Unified Foundation Model for Infrared and Visible Modalities Unified perceptual parsing for scene understand- ing

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T16:32:45.395213Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-18T16:28:48.494939Z digest=sha256:b5af7e44e2966ee863ffd8606c6b6c7567c19862e5f79bd14934afac3044fb9f

Observation 0a8b5769-fd20-45b2-a733-9da959ddd5e7 · outbound

This paper cites Hitea: Hierarchical temporal- aware video-language pre-training.

UNIV: Unified Foundation Model for Infrared and Visible Modalities Hitea: Hierarchical temporal- aware video-language pre-training

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T16:32:45.360978Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-18T16:28:48.494939Z digest=sha256:18d950cf9b7889eb247bca10e70164b9b26a435dbcc185dd85ee69e66c130678

Observation 36226bf4-198f-49f8-b791-891a9717c59d · outbound

This paper cites Sigmoid loss for language image pre-training.

UNIV: Unified Foundation Model for Infrared and Visible Modalities Sigmoid loss for language image pre-training

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T16:32:45.348380Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-18T16:28:48.494939Z digest=sha256:174d4e80ae410bb3b4f3673777939c7502c9f330a014a00c9421b5443741221a

Observation 2f978e6a-9ae7-466f-9d25-219877188654 · outbound

This paper cites PAD: Self-Supervised Pre-Training with Patchwise-Scale Adapter for Infrared Images.

UNIV: Unified Foundation Model for Infrared and Visible Modalities PAD: Self-Supervised Pre-Training with Patchwise-Scale Adapter for Infrared Images

Reference 42

Resolution
verified exact
arxiv_id, observed 2026-05-18T16:31:37.157492Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-18T16:28:48.494939Z digest=sha256:7629bfdb8bd6afa6ed53c2c022ed33707a4c09a9029db8351484cf39cf9e9d63

Observation deedfaa1-d2b8-4a73-a7ab-be4fab009306 · outbound

This paper cites UNIP: Rethinking Pre-trained Attention Patterns for Infrared Semantic Segmentation.

UNIV: Unified Foundation Model for Infrared and Visible Modalities UNIP: Rethinking Pre-trained Attention Patterns for Infrared Semantic Segmentation

Reference 43

Resolution
verified exact
arxiv_id, observed 2026-05-18T16:31:37.166147Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-18T16:28:48.494939Z digest=sha256:fb64327b80d9d52c03fc5e3dd558209d91e268f37e7bc3360372389e68541b6f

Observation 5c14ab0c-542b-4b1d-af81-d877ec393311 · outbound

This paper cites Scene parsing through ade20k dataset.

UNIV: Unified Foundation Model for Infrared and Visible Modalities Scene parsing through ade20k dataset

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T16:32:45.381215Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-18T16:28:48.494939Z digest=sha256:a65890e81442c1fbf4c039dd7257c3078a55e12ed6adb4e586d2ddc5ddfd7a49

Observation fb6eea95-5fe6-4981-b522-70b61f53c47b · outbound

This paper cites iBOT: Image BERT Pre-Training with Online Tokenizer.

UNIV: Unified Foundation Model for Infrared and Visible Modalities iBOT: Image BERT Pre-Training with Online Tokenizer

Reference 45

Resolution
verified exact
local_arxiv, observed 2026-05-18T16:31:37.175410Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-18T16:28:48.494939Z digest=sha256:bf7389420495ebd74281cd559b1f0e89729817620343148ef83061a17b527bb1

Pith citing papers

Observation 5d2a49d0-3772-4110-b706-4cb64ef106b9 · inbound

Not All Patches are Equal: Sampling Matters for Visible-Infrared Pre-Training cites this paper.

Not All Patches are Equal: Sampling Matters for Visible-Infrared Pre-Training UNIV: Unified Foundation Model for Infrared and Visible Modalities

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-01T10:27:20.578105Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T10:27:20.578105Z digest=sha256:317350598d42075e5e41366ac494c72a03ccd5428265a51262aeaaed5c435824