Abstract:Few-shot learning presents a critical solution for cancer diagnosis in computational pathology (CPath), addressing fundamental limitations in data availability, particularly the scarcity of expert annotations and patient privacy constraints. A key challenge in this paradigm stems from the inherent disparity between the limited training set of whole slide images (WSIs) and the enormous number of contained patches, where a significant portion of these patches lacks diagnostically relevant information, potentially diluting the model's ability to learn and focus on critical diagnostic features. While recent works attempt to address this by incorporating additional knowledge, several crucial gaps hinder further progress: (1) despite the emergence of powerful pathology foundation models (FMs), their potential remains largely untapped, with most approaches limiting their use to basic feature extraction; (2) current language guidance mechanisms attempt to align text prompts with vast numbers of WSI patches all at once, struggling to leverage rich pathological semantic information. To this end, we introduce the knowledge-enhanced adaptive visual compression framework, dubbed FOCUS, which uniquely combines pathology FMs with language prior knowledge to enable a focused analysis of diagnostically relevant regions by prioritizing discriminative WSI patches. Our approach implements a progressive three-stage compression strategy: we first leverage FMs for global visual redundancy elimination, and integrate compressed features with language prompts for semantic relevance assessment, then perform neighbor-aware visual token filtering while preserving spatial coherence. Extensive experiments on pathological datasets spanning breast, lung, and ovarian cancers demonstrate its superior performance in few-shot pathology diagnosis. Code will be made available at <a class="link-external link-https" href="https://github.com/dddavid4real/FOCUS" rel="external noopener nofollow">this https URL</a>.

What problem does this paper attempt to address?

The problem that this paper attempts to solve is how to use limited training samples for effective whole - slide image (WSI) classification in computational pathology (CPath), especially in the field of cancer diagnosis. Specifically, the paper focuses on the following key challenges: 1. **Data Scarcity**: In computational pathology, due to the scarcity of expert - annotated data and the limitations of patient privacy, the amount of data available for training is very limited. This makes it difficult to apply traditional deep - learning methods because they usually require a large amount of annotated data to train models. 2. **Redundant Information**: WSIs usually contain a large number of regions with no diagnostic value, and these regions will dilute the model's ability to learn key diagnostic features. Therefore, how to effectively identify and focus on regions with diagnostic value has become an important issue. 3. **Under - exploited Potential of Base Models**: Although existing pathological base models (FMs) have demonstrated strong capabilities, most methods only use them for basic feature extraction and do not fully utilize their rich semantic understanding and context - learning capabilities. 4. **Limitations of Language - Guided Mechanisms**: Current language - guided mechanisms attempt to align text prompts with a large number of WSI patches all at once, but this approach makes it difficult to fully utilize complex pathological semantic information, especially when only a few patches contain diagnosis - related patterns. To address the above challenges, the paper proposes a knowledge - enhanced adaptive visual compression framework - FOCUS. This framework combines pathological base models and language prior knowledge to achieve focused analysis of diagnosis - related regions, thereby improving the WSI classification performance in the case of a small number of samples.

FOCUS: Knowledge-enhanced Adaptive Visual Compression for Few-shot Whole Slide Image Classification

Rapid Whole Slide Imaging via Learning-based Two-shot Virtual Autofocusing

Pathology-knowledge Enhanced Multi-instance Prompt Learning for Few-shot Whole Slide Image Classification

Focus Quality Assessment of High-Throughput Whole Slide Imaging in Digital Pathology

Multi-Cohort Framework with Cohort-Aware Attention and Adversarial Mutual-Information Minimization for Whole Slide Image Classification

Data-efficient and weakly supervised computational pathology on whole-slide images

Semantics-Aware Attention Guidance for Diagnosing Whole Slide Images

Attention-based whole-slide image compression achieves pathologist-level pre-screening of multi-organ routine histopathology biopsies

Focus on Focus: Focus-oriented Representation Learning and Multi-view Cross-modal Alignment for Glioma Grading

A Knowledge-enhanced Pathology Vision-language Foundation Model for Cancer Diagnosis

Federated learning for computational pathology on gigapixel whole slide images

Automatic Whole Slide Pathology Image Diagnosis Framework Via Unit Stochastic Selection and Attention Fusion

Cluster-to-Conquer: A Framework for End-to-End Multi-Instance Learning for Whole Slide Image Classification

The Rise of AI Language Pathologists: Exploring Two-level Prompt Learning for Few-shot Weakly-supervised Whole Slide Image Classification

Learning to autofocus in whole slide imaging via physics-guided deep cascade networks

Overcoming the limitations of patch-based learning to detect cancer in whole slide images

PATHS: A Hierarchical Transformer for Efficient Whole Slide Image Analysis

FFusionCGAN: An end-to-end fusion method for few-focus images using conditional GAN in cytopathological digital slides

Clinical-grade computational pathology using weakly supervised deep learning on whole slide images

Robust whole slide image analysis for cervical cancer screening using deep learning

Contrastive learning-based histopathological features infer molecular subtypes and clinical outcomes of breast cancer from unannotated whole slide images