Training Like a Medical Resident: Context-Prior Learning Toward Universal Medical Image Segmentation

Yunhe Gao,Zhuowei Li,Di Liu,Mu Zhou,Shaoting Zhang,Dimitris N. Metaxas

2024-04-07

Abstract:A major focus of clinical imaging workflow is disease diagnosis and management, leading to medical imaging datasets strongly tied to specific clinical objectives. This scenario has led to the prevailing practice of developing task-specific segmentation models, without gaining insights from widespread imaging cohorts. Inspired by the training program of medical radiology residents, we propose a shift towards universal medical image segmentation, a paradigm aiming to build medical image understanding foundation models by leveraging the diversity and commonality across clinical targets, body regions, and imaging modalities. Towards this goal, we develop Hermes, a novel context-prior learning approach to address the challenges of data heterogeneity and annotation differences in medical image segmentation. In a large collection of eleven diverse datasets (2,438 3D images) across five modalities (CT, PET, T1, T2 and cine MRI) and multiple body regions, we demonstrate the merit of the universal paradigm over the traditional paradigm on addressing multiple tasks within a single model. By exploiting the synergy across tasks, Hermes achieves state-of-the-art performance on all testing datasets and shows superior model scalability. Results on two additional datasets reveals Hermes' strong performance for transfer learning, incremental learning, and generalization to downstream tasks. Hermes's learned priors demonstrate an appealing trait to reflect the intricate relations among tasks and modalities, which aligns with the established anatomical and imaging principles in radiology. The code is available: <a class="link-external link-https" href="https://github.com/yhygao/universal-medical-image-segmentation" rel="external noopener nofollow">this https URL</a>.

Computer Vision and Pattern Recognition

What problem does this paper attempt to address?

The paper proposes a new medical image segmentation method called Hermes to address the limitations of existing task-specific models. The current practice is to develop separate segmentation models for each medical target or imaging modality, which limits the model's generalization ability and the insights from a wide imaging queue. Hermes achieves universal medical image segmentation by learning contextual prior knowledge, handling data heterogeneity and annotation discrepancies. It aims to build a foundational model capable of understanding different clinical targets, body regions, and imaging modalities. Experimental results demonstrate that Hermes performs well on multiple tasks, exhibiting better model scalability and strong performance in transfer learning, incremental learning, and downstream task generalization.

Training Like a Medical Resident: Context-Prior Learning Toward Universal Medical Image Segmentation

Super-Resolution Based Patch-Free 3D Medical Image Segmentation with Self-Supervised Guidance

Patch-Free 3D Medical Image Segmentation Driven by Super-Resolution Technique and Self-Supervised Guidance

Domain Adaptation Meets Zero-Shot Learning: an Annotation-Efficient Approach to Multi-Modality Medical Image Segmentation

Many Birds, One Stone: Medical Image Segmentation with Multiple Partially Labeled Datasets

UniSeg: A Prompt-driven Universal Segmentation Model as well as A Strong Representation Learner

Universal Model for 3D Medical Image Analysis

MedContext: Learning Contextual Cues for Efficient Volumetric Medical Segmentation

MultiTalent: A Multi-Dataset Approach to Medical Image Segmentation

UniverSeg: Universal Medical Image Segmentation

Tri-Directional Tasks Complementary Learning for Unsupervised Domain Adaptation of Cross-modality Medical Image Semantic Segmentation.

Correlative studies of cell wall enzymes and growth.

Unified Medical Image Pre-training in Language-Guided Common Semantic Space

UniMOS: A Universal Framework For Multi-Organ Segmentation Over Label-Constrained Datasets

A New Framework of Swarm Learning Consolidating Knowledge from Multi-Center Non-IID Data for Medical Image Segmentation

MulModSeg: Enhancing Unpaired Multi-Modal Medical Image Segmentation with Modality-Conditioned Text Embedding and Alternating Training

A Dimension Hybrid Framework for Multimodal Medical Image Segmentation

Unleashing the Potential of Vision-Language Pre-Training for 3D Zero-Shot Lesion Segmentation via Mask-Attribute Alignment

Multi-Level Global Context Cross Consistency Model for Semi-Supervised Ultrasound Image Segmentation with Diffusion Model

One Network to Segment Them All: A General, Lightweight System for Accurate 3D Medical Image Segmentation

MedUniSeg: 2D and 3D Medical Image Segmentation via a Prompt-driven Universal Model