A local-global transformer-based model for person re-identification.
basic_science · Level V
Where this comes from
- Record sourced from PubMed, PMID 41231884.
- Also identified by DOI 10.1371/journal.pone.0335848 and PMC identifier 12614614.
- No licence information is recorded for this record.
- Because redistribution is not established, this page shows the abstract only. Follow the links below for the full text.
Abstract
Person re-identification (ReID) aims to recognize a specific individual across various camera views. State-of-the-art methods have shown that both Transformer-based and CNN-based methods deliver competitive performance. However, Transformer-based methods tend to overlook local features, as they primarily process input sequences holistically, rather than focusing on individual elements or small groups within the sequence. To address this limitation, we introduce an innovative Transformer-based person ReID model that effectively integrates local and global features. The Local Attention Module is added to capture fine-grained features, which are then combined with global features to enhance the model's recognition accuracy. Given the importance of positional information in image data, relative position encoding is incorporated within the Local Attention Module. This encoding method better captures the relative positional relationships between different tokens in an image, thereby improving the model's comprehension of the structural information of the image. Experimental results indicate that the Rank-1 of our model respectively improves by 0.7% and 0.9% on the Market-1501 and DukeMTMC-reID benchmark datasets for person ReID.
Medical subject headings
- Biometric Identification
- Image Processing, Computer-Assisted