Abstract:Gait recognition has attracted increasing attention from academia and industry as a human recognition technology from a distance in non-intrusive ways without requiring cooperation. Although advanced methods have achieved impressive success in lab scenarios, most of them perform poorly in the wild. Recently, some Convolution Neural Networks (ConvNets) based methods have been proposed to address the issue of gait recognition in the wild. However, the temporal receptive field obtained by convolution operations is limited for long gait sequences. If directly replacing convolution blocks with visual transformer blocks, the model may not enhance a local temporal receptive field, which is important for covering a complete gait cycle. To address this issue, we design a Global-Local Temporal Receptive Field Network (GLGait). GLGait employs a Global-Local Temporal Module (GLTM) to establish a global-local temporal receptive field, which mainly consists of a Pseudo Global Temporal Self-Attention (PGTA) and a temporal convolution operation. Specifically, PGTA is used to obtain a pseudo global temporal receptive field with less memory and computation complexity compared with a multi-head self-attention (MHSA). The temporal convolution operation is used to enhance the local temporal receptive field. Besides, it can also aggregate pseudo global temporal receptive field to a true holistic temporal receptive field. Furthermore, we also propose a Center-Augmented Triplet Loss (CTL) in GLGait to reduce the intra-class distance and expand the positive samples in the training stage. Extensive experiments show that our method obtains state-of-the-art results on in-the-wild datasets, $i.e.$, Gait3D and GREW. The code is available at <a class="link-external link-https" href="https://github.com/bgdpgz/GLGait" rel="external noopener nofollow">this https URL</a>.

Learning Visual Prompt for Gait Recognition

GLGait: A Global-Local Temporal Receptive Field Network for Gait Recognition in the Wild

Human Gait Recognition Based on Self-Adaptive Hidden Markov Model

Human Gait Recognition Based on Frame-by-Frame Gait Energy Images and Convolutional Long Short-Term Memory

DyGait: Exploiting Dynamic Representations for High-performance Gait Recognition

Dynamic Aggregated Network for Gait Recognition

BigGait: Learning Gait Representation You Want by Large Vision Models

GaitMPL: Gait Recognition with Memory-Augmented Progressive Learning

DeepGait: A Learning Deep Convolutional Representation for Gait Recognition

GaitCTCG: cross-view gait recognition via cascaded residual temporal shift and comprehensive multi-granularity learning

Multi-scale Context-aware Network with Transformer for Gait Recognition

On Learning Disentangled Representations for Gait Recognition

GaitContour: Efficient Gait Recognition based on a Contour-Pose Representation

Human Gait Recognition Based on Frontal-View Walking Sequences Using Multi-modal Feature Representations and Learning

Context-Sensitive Temporal Feature Learning for Gait Recognition

Exploring Deep Models for Practical Gait Recognition

TAG: A Temporal Attentive Gait Network for Cross-View Gait Recognition

GaitSet: Cross-view Gait Recognition through Utilizing Gait as a Deep Set

Gait Recognition in the Wild with Multi-hop Temporal Switch

TAG: A Temporal AttentiveGait Network for Cross-View Gait Recognition

GaitTAKE: Gait Recognition by Temporal Attention and Keypoint-guided Embedding