REVIEW 3 cited by
Multi-Memory Matching for Unsupervised Visible-Infrared Person Re-Identification
Not yet reviewed by Pith; the record is open.
This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.
SPECIMEN: schema-true, not a live event
T0 review · schema-true
One-sentence machine reading of the paper's core claim.
pith:XXXXXXXX · record.json · timestamp
read the original abstract
Unsupervised visible-infrared person re-identification (USL-VI-ReID) is a promising yet challenging retrieval task. The key challenges in USL-VI-ReID are to effectively generate pseudo-labels and establish pseudo-label correspondences across modalities without relying on any prior annotations. Recently, clustered pseudo-label methods have gained more attention in USL-VI-ReID. However, previous methods fell short of fully exploiting the individual nuances, as they simply utilized a single memory that represented an identity to establish cross-modality correspondences, resulting in ambiguous cross-modality correspondences. To address the problem, we propose a Multi-Memory Matching (MMM) framework for USL-VI-ReID. We first design a Cross-Modality Clustering (CMC) module to generate the pseudo-labels through clustering together both two modality samples. To associate cross-modality clustered pseudo-labels, we design a Multi-Memory Learning and Matching (MMLM) module, ensuring that optimization explicitly focuses on the nuances of individual perspectives and establishes reliable cross-modality correspondences. Finally, we design a Soft Cluster-level Alignment (SCA) module to narrow the modality gap while mitigating the effect of noise pseudo-labels through a soft many-to-many alignment strategy. Extensive experiments on the public SYSU-MM01 and RegDB datasets demonstrate the reliability of the established cross-modality correspondences and the effectiveness of our MMM. The source codes will be released.
Forward citations
Cited by 3 Pith papers
-
MambaPro: Multi-Modal Object Re-Identification with Mamba Aggregation and Synergistic Prompt
MambaPro combines CLIP with a parallel adapter, synergistic residual prompts, and Mamba aggregation to achieve state-of-the-art mAP on RGBNT201, RGBNT100, and MSVR310.
-
DeMo: Decoupled Feature-Based Mixture of Experts for Multi-Modal Object Re-Identification
DeMo improves multi-modal object re-identification by decoupling RGB, NIR, and TIR features into seven attention-derived streams and weighting them with an attention-triggered mixture of experts.
-
Unity is Strength: Unifying Convolutional and Transformeral Features for Better Person Re-Identification
FusionReID, a dual-branch CNN-Transformer network with mutual cross-attention fusion, achieves state-of-the-art person re-identification on Market1501, DukeMTMC, and MSMT17.
Discussion (0). Continue with ORCID to comment.