Long-Tailed Effect Study in Remote Sensing Semantic Segmentation Based on Graph Kernel Principles

Wei Cui,Zhanyun Feng,Jiale Chen,Xing Xu,Yueling Tian,Huilin Zhao,Chenglei Wang,Cui,Feng,Chen,Xu,Tian,Zhao,Wang

DOI: https://doi.org/10.3390/rs16081398

IF: 5

2024-04-16

Remote Sensing

Abstract:The performance of semantic segmentation in remote sensing, based on deep learning models, depends on the training data. A commonly encountered issue is the imbalanced long-tailed distribution of data, where the head classes contain the majority of samples while the tail classes have fewer samples. When training with long-tailed data, the head classes dominate the training process, resulting in poorer performance in the tail classes. To address this issue, various strategies have been proposed, such as resampling, reweighting, and transfer learning. However, common resampling methods suffer from overfitting to the tail classes while underfitting the head classes, and reweighting methods are limited in the extreme imbalanced case. Additionally, transfer learning tends to transfer patterns learned from the head classes to the tail classes without rigorously validating its generalizability. These methods often lack additional information to assist in the recognition of tail class objects, thus limiting performance improvements and constraining generalization ability. To tackle the abovementioned issues, a graph neural network based on the graph kernel principle is proposed for the first time. By leveraging the graph kernel, structural information for tail class objects is obtained, serving as additional contextual information beyond basic visual features. This method partially compensates for the imbalance between tail and head class object information without compromising the recognition accuracy of head classes objects. The experimental results demonstrate that this study effectively addresses the poor recognition performance of small and rare targets, partially alleviates the issue of spectral confusion, and enhances the model's generalization ability.

environmental sciences,imaging science & photographic technology,remote sensing,geosciences, multidisciplinary

What problem does this paper attempt to address?

The paper attempts to address the issue of poor recognition performance of tail classes (i.e., categories with fewer samples) in remote sensing semantic segmentation tasks due to the long-tail effect of data distribution. Specifically, the paper points out that in deep learning-based remote sensing semantic segmentation, the imbalanced distribution of training data is a common problem, where the number of samples in head classes is large, while the number of samples in tail classes is small. This imbalance can cause the model to overly focus on head classes during the training process, thereby affecting the recognition performance of tail classes. To tackle this challenge, the paper proposes a graph neural network method based on the principle of graph kernels. This method enhances the feature representation of tail classes by extracting the graph structure information of tail class objects and their neighboring nodes as supplementary information, thereby partially compensating for the information imbalance between head and tail classes. Experimental results show that this method effectively improves the recognition performance of small and rare targets, partially alleviates the spectral confusion problem, and enhances the generalization ability of the model.

Long-Tailed Effect Study in Remote Sensing Semantic Segmentation Based on Graph Kernel Principles

Remote Sensing Scene Graph and Knowledge Graph Matching with Parallel Walking Algorithm

On Size-Oriented Long-Tailed Graph Classification of Graph Neural Networks

Optimizing Spatial Relationships in GCN to Improve the Classification Accuracy of Remote Sensing Images

Fusion of hierarchical class graphs for remote sensing semantic segmentation

Knowledge and Geo-Object Based Graph Convolutional Network for Remote Sensing Semantic Segmentation

Knowledge and Spatial Pyramid Distance-Based Gated Graph Attention Network for Remote Sensing Semantic Segmentation

Long-Tailed Object Detection for Multimodal Remote Sensing Images

DECOR: Dynamic Decoupling and Multi-Objective Optimization for Long-tailed Remote Sensing Image Classification

Graph Information Bottleneck for Remote Sensing Segmentation

A Stage-Adaptive Selective Network with Position Awareness for Semantic Segmentation of LULC Remote Sensing Images

Long-tailed Visual Recognition with Deep Models: A Methodological Survey and Evaluation

Advancing high-resolution remote sensing: a compact and powerful approach to semantic segmentation

A multiscale bidirectional fuzzy-driven learning network for remote sensing image segmentation

Remote sensing image semantic segmentation method based on small target and edge feature enhancement

Transformer-induced graph reasoning for multimodal semantic segmentation in remote sensing

Center-Wise Feature Consistency Learning for Long-Tailed Remote Sensing Object Recognition

DECOR: Dynamic Decoupling and Multiobjective Optimization for Long-Tailed Remote Sensing Image Classification

Towards Long-Tailed Recognition for Graph Classification Via Collaborative Experts

Text-Guided Diverse Image Synthesis for Long-Tailed Remote Sensing Object Classification

Fuzzy neighbourhood neural network for high-resolution remote sensing image segmentation