Abstract:Facial expression data is characterized by a significant imbalance, with most collected data showing happy or neutral expressions and fewer instances of fear or disgust. This imbalance poses challenges to facial expression recognition (FER) models, hindering their ability to fully understand various human emotional states. Existing FER methods typically report overall accuracy on highly imbalanced test sets but exhibit low performance in terms of the mean accuracy across all expression classes. In this paper, our aim is to address the imbalanced FER problem. Existing methods primarily focus on learning knowledge of minor classes solely from minor-class samples. However, we propose a novel approach to extract extra knowledge related to the minor classes from both major and minor class samples. Our motivation stems from the belief that FER resembles a distribution learning task, wherein a sample may contain information about multiple classes. For instance, a sample from the major class surprise might also contain useful features of the minor class fear. Inspired by that, we propose a novel method that leverages re-balanced attention maps to regularize the model, enabling it to extract transformation invariant information about the minor classes from all training samples. Additionally, we introduce re-balanced smooth labels to regulate the cross-entropy loss, guiding the model to pay more attention to the minor classes by utilizing the extra information regarding the label distribution of the imbalanced training data. Extensive experiments on different datasets and backbones show that the two proposed modules work together to regularize the model and achieve state-of-the-art performance under the imbalanced FER task. Code is available at <a class="link-external link-https" href="https://github.com/zyh-uaiaaaa" rel="external noopener nofollow">this https URL</a>.

Learning to Amend Facial Expression Representation via De-albino and Affinity

DR-FER: Discriminative and Robust Representation Learning for Facial Expression Recognition

An Improved SimAM Based CNN for Facial Expression Recognition

Efficient Facial Expression Recognition with Representation Reinforcement Network and Transfer Self-Training for Human–Machine Interaction

Facial expression recognition through multi-level features extraction and fusion

Facial Expression Recognition with Contrastive Learning and Uncertainty-Guided Relabeling

Enhanced Dual-Level Representations for Facial Expression Recognition

Adaptive multilayer perceptual attention network for facial expression recognition

Evaluation and analysis of visual perception using attention-enhanced computation in multimedia affective computing

Improved facial emotion recognition model based on a novel deep convolutional structure

PAtt-Lite: Lightweight Patch and Attention MobileNet for Challenging Facial Expression Recognition

Patch-Aware Representation Learning for Facial Expression Recognition

FERMixNet: an Occlusion Robust Facial Expression Recognition Model with Facial Mixing Augmentation and Mid-Level Representation Learning

Adaptive Deep Disturbance-Disentangled Learning for Facial Expression Recognition

Real Emotion Seeker: Recalibrating Annotation for Facial Expression Recognition

Visual Scene-Aware Hybrid and Multi-Modal Feature Aggregation for Facial Expression Recognition

Leave No Stone Unturned: Mine Extra Knowledge for Imbalanced Facial Expression Recognition

Multi-Head Attention Affinity Diversity Sharing Network for Facial Expression Recognition

Facial Expression Recognition Using Hybrid Features of Pixel and Geometry

AU-Oriented Expression Decomposition Learning for Facial Expression Recognition.

Multi Loss-based Feature Fusion and Top Two Voting Ensemble Decision Strategy for Facial Expression Recognition in the Wild