Abstract:Deep learning has gradually become powerful in segmenting and classifying aerial images. However, in remote sensing applications, the lack of training datasets and the difficulty of accuracy assessment have always been challenges for the deep learning based classification. In recent years, interactive semantic segmentation proposed in computer vision has achieved an ideal state of human-computer interaction segmentation. It can provide expert experience and utilize deep learning for efficient segmentation. However, few papers discussed its application in remote sensing imagery. This study aims to bridge the gap between interactive segmentation and remote sensing analysis by conducting a benchmark study on various interactive segmentation models. We assessed the performance of five state-of-the-art interactive segmentation methods (Reviving Iterative Training with Mask Guidance for Interactive Segmentation (RITM), FocalClick, SimpleClick, Iterative Click Loss (ICL), and Segment Anything (SAM)) on two high-resolution aerial imagery datasets. The Cascade-Forward Refinement approach, an innovative inference strategy for interactive segmentation, was also introduced to enhance the segmentation results. We evaluated these methods on various land cover types, object sizes, and band combinations in the datasets. SimpleClick model consistently outperformed the other methods in our experiments. Conversely, the SAM performed less effectively than other models. Building upon these findings, we developed an online tool called RSISeg for interactive segmentation of remote sensing data. RSISeg incorporates a well-performing interactive model that is finetuned with remote sensing data. Compared to existing interactive segmentation tools, RSISeg offers robust interactivity, modifiability, and adaptability to remote sensing data.

A Click-Based Interactive Segmentation Network for Point Clouds

One-Click-Based Perception for Interactive Image Segmentation

FocalClick: Towards Practical Interactive Image Segmentation.

3D Object Segmentation Using Cross-Window Point Transformer with Latent Semantic Boundary Guidance

Refining Segmentation On-the-Fly: An Interactive Framework for Point Cloud Semantic Segmentation

Interactive Object Segmentation in 3D Point Clouds

PiClick: Picking the desired mask from multiple candidates in click-based interactive segmentation

AGILE3D: Attention Guided Interactive Multi-object 3D Segmentation

Associate Semantic-Instance Segmentation of 3D Point Clouds Based on Local Feature Extraction

CSANet: Cross-self attention guided by semantic click embedding for interactive segmentation

Scale Disparity of Instances in Interactive Point Cloud Segmentation

Interactive segmentation in aerial images: a new benchmark and an open access web-based tool

Multi-Scale Point-Wise Convolutional Neural Networks for 3D Object Segmentation From LiDAR Point Clouds in Large-Scale Environments

ClickAttention: Click Region Similarity Guided Interactive Segmentation

PseudoClick: Interactive Image Segmentation with Click Imitation

RClicks: Realistic Click Simulation for Benchmarking Interactive Segmentation

Aircraft Segmentation Based On Deep Learning Framework : From Extreme Points To Remote Sensing Image Segmentation

CPSeg: Cluster-free Panoptic Segmentation of 3D LiDAR Point Clouds

ClickSeg: 3D Instance Segmentation with Click-Level Weak Annotations