Pose-guided spatial alignment and key frame selection for one-shot video-based person re-identification

N/ACitations
Citations of this article
18Readers
Mendeley users who have this article in their library.

This article is free to access.

Abstract

One-shot video-based person re-identification exploits the unlabeled data by using a single-labeled sample for each individual to train a model and to reduce the need for laborious labeling. Although recent works focusing on this task have made some achievements, most state-of-the-art models are vulnerable to misalignment, pose variation and corrupted frames. To address these challenges, we propose a one-shot video-based person re-identification model based on pose-guided spatial alignment and KFS. First, a spatial transformer sub-network trained using pose-guided regression is employed to perform the spatial alignment. Second, we propose a novel training strategy based on KFS. Key frames with abruptly changing poses are deliberately identified and selected to make the network adaptive to pose variation. Finally, we propose a frame feature pooling method by incorporating long short-term memory with an attention mechanism to reduce the influence of corrupted frames. Comprehensive experiments are presented based on the MARS and DukeMTMC-VideoReID datasets. The mAP values for these datasets reach 46.5% and 68.4%, respectively, demonstrating that the proposed model achieves significant improvements over state-of-the-art one-shot person re-identification methods.

Cite

CITATION STYLE

APA

Chen, Y., Huang, T., Niu, Y., Ke, X., & Lin, Y. (2019). Pose-guided spatial alignment and key frame selection for one-shot video-based person re-identification. IEEE Access, 7, 78991–79004. https://doi.org/10.1109/ACCESS.2019.2922679

Register to see more suggestions

Mendeley helps you to discover research relevant for your work.

Already have an account?

Save time finding and organizing research with Mendeley

Sign up for free