Search papers, labs, and topics across Lattice.
This paper introduces Structural-Semantic Reciprocal Learning (SSRL) to tackle the challenges of unsupervised visible-infrared person re-identification (USVI-ReID) by transforming the traditional open-loop association into a self-correcting closed-loop system. The method employs Fine-grained Structural Decoupling (FSD) to derive discriminative body-part primitives, enhancing spatial consistency, while the Closed-loop Semantic Calibration (CSC) mechanism refines pseudo-labels by reconstructing shared semantic prototypes throughout training. Experimental results show that SSRL outperforms state-of-the-art USVI-ReID methods on benchmark datasets, achieving competitive results even against several supervised approaches.
SSRL not only bridges the modality gap in USVI-ReID but also filters out pseudo-label noise, leading to superior performance compared to existing methods.
Unsupervised visible-infrared person re-identification (USVI-ReID) is challenging due to the large modality gap and the lack of cross-modal identity annotations. Progressive association paradigms have been proposed to gradually bridge the gap, but they suffer from two critical bottlenecks: reliance on ambiguous global representations and unchecked propagation of pseudo-label noise in an open-loop manner. To address these issues, we propose Structural-Semantic Reciprocal Learning (SSRL), a framework that transforms open-loop association into a self-correcting closed-loop system. Structurally, we introduce Fine-grained Structural Decoupling (FSD) to extract discriminative body-part primitives as reliable spatial anchors, complementing ambiguous holistic silhouettes with spatially consistent structural details. Semantically, we design a Closed-loop Semantic Calibration (CSC) mechanism that reconstructs shared semantic prototypes at each epoch and feeds them back into the training loop, effectively filtering pseudo-label noise before the next clustering cycle. Through the reciprocal interaction between structural and semantic learning, SSRL achieves robust cross-modal representation. Extensive experiments demonstrate the competitive performance of SSRL against state-of-the-art USVI-ReID methods on both SYSU-MM01 and RegDB, notably surpassing several supervised counterparts on RegDB.