Abstract:The Segment Anything Model (SAM) achieves remarkable promptable segmentation given high-quality prompts which, however, often require good skills to specify. To make SAM robust to casual prompts, this paper presents the first comprehensive analysis on SAM's segmentation stability across a diverse spectrum of prompt qualities, notably imprecise bounding boxes and insufficient points. Our key finding reveals that given such low-quality prompts, SAM's mask decoder tends to activate image features that are biased towards the background or confined to specific object parts. To mitigate this issue, our key idea consists of calibrating solely SAM's mask attention by adjusting the sampling locations and amplitudes of image features, while the original SAM model architecture and weights remain unchanged. Consequently, our deformable sampling plugin (DSP) enables SAM to adaptively shift attention to the prompted target regions in a data-driven manner, facilitated by our effective robust training strategy (RTS). During inference, dynamic routing plugin (DRP) is proposed that toggles SAM between the deformable and regular grid sampling modes, conditioned on the input prompt quality. Thus, our solution, termed Stable-SAM, offers several advantages: 1) improved SAM's segmentation stability across a wide range of prompt qualities, while 2) retaining SAM's powerful promptable segmentation efficiency and generality, with 3) minimal learnable parameters (0.08 M) and fast adaptation (by 1 training epoch). Extensive experiments across multiple datasets validate the effectiveness and advantages of our approach, underscoring Stable-SAM as a more robust solution for segmenting anything. Codes will be released upon acceptance. <a class="link-external link-https" href="https://github.com/fanq15/Stable-SAM" rel="external noopener nofollow">this https URL</a>

Char-SAM: Turning Segment Anything Model into Scene Text Segmentation Annotator with Character-level Visual Prompts

AM-SAM: Automated Prompting and Mask Calibration for Segment Anything Model

AI-SAM: Automatic and Interactive Segment Anything Model

SAM Fails to Segment Anything? – SAM-Adapter: Adapting SAM in Underperformed Scenes: Camouflage, Shadow, Medical Image Segmentation, and More

Stable Segment Anything Model

PA-SAM: Prompt Adapter SAM for High-Quality Image Segmentation

Hi-SAM: Marrying Segment Anything Model for Hierarchical Text Segmentation

SAM-Adapter: Adapting Segment Anything in Underperformed Scenes

SAM-SP: Self-Prompting Makes SAM Great Again

SAM-CP: Marrying SAM with Composable Prompts for Versatile Segmentation

All-in-SAM: from Weak Annotation to Pixel-wise Nuclei Segmentation with Prompt-based Finetuning

BioSAM: Generating SAM Prompts From Superpixel Graph for Biological Instance Segmentation

MaskSAM: Towards Auto-prompt SAM with Mask Classification for Medical Image Segmentation

SAM-RSIS: Progressively Adapting SAM With Box Prompting to Remote Sensing Image Instance Segmentation

SAM-Path: A Segment Anything Model for Semantic Segmentation in Digital Pathology

Segment Anything without Supervision

RSAM-Seg: A SAM-based Approach with Prior Knowledge Integration for Remote Sensing Image Semantic Segmentation

Robust Box Prompt based SAM for Medical Image Segmentation

AGSAM: Agent-Guided Segment Anything Model for Automatic Segmentation in Few-Shot Scenarios

A Survey on Segment Anything Model (SAM): Vision Foundation Model Meets Prompt Engineering

CoSAM: Self-Correcting SAM for Domain Generalization in 2D Medical Image Segmentation