Get our free extension to see links to code for papers anywhere online!Free add-on: code for papers everywhere!Free add-on: See code for papers anywhere!

Add to Chrome

Add to Firefox

Add to Edge

Title:GLGait: A Global-Local Temporal Receptive Field Network for Gait Recognition in the Wild

Aug 13, 2024

Guozhen Peng, Yunhong Wang, Yuwei Zhao, Shaoxiong Zhang, Annan Li

Figure 1 for GLGait: A Global-Local Temporal Receptive Field Network for Gait Recognition in the Wild

Figure 2 for GLGait: A Global-Local Temporal Receptive Field Network for Gait Recognition in the Wild

Figure 3 for GLGait: A Global-Local Temporal Receptive Field Network for Gait Recognition in the Wild

Figure 4 for GLGait: A Global-Local Temporal Receptive Field Network for Gait Recognition in the Wild

Share this with someone who'll enjoy it:

Abstract:Gait recognition has attracted increasing attention from academia and industry as a human recognition technology from a distance in non-intrusive ways without requiring cooperation. Although advanced methods have achieved impressive success in lab scenarios, most of them perform poorly in the wild. Recently, some Convolution Neural Networks (ConvNets) based methods have been proposed to address the issue of gait recognition in the wild. However, the temporal receptive field obtained by convolution operations is limited for long gait sequences. If directly replacing convolution blocks with visual transformer blocks, the model may not enhance a local temporal receptive field, which is important for covering a complete gait cycle. To address this issue, we design a Global-Local Temporal Receptive Field Network (GLGait). GLGait employs a Global-Local Temporal Module (GLTM) to establish a global-local temporal receptive field, which mainly consists of a Pseudo Global Temporal Self-Attention (PGTA) and a temporal convolution operation. Specifically, PGTA is used to obtain a pseudo global temporal receptive field with less memory and computation complexity compared with a multi-head self-attention (MHSA). The temporal convolution operation is used to enhance the local temporal receptive field. Besides, it can also aggregate pseudo global temporal receptive field to a true holistic temporal receptive field. Furthermore, we also propose a Center-Augmented Triplet Loss (CTL) in GLGait to reduce the intra-class distance and expand the positive samples in the training stage. Extensive experiments show that our method obtains state-of-the-art results on in-the-wild datasets, $i.e.$, Gait3D and GREW. The code is available at https://github.com/bgdpgz/GLGait.

* Accepted by ACM MM2024

View paper on

Share this with someone who'll enjoy it:

Title:GLGait: A Global-Local Temporal Receptive Field Network for Gait Recognition in the Wild

Paper and Code