Abstract:Dermatological conditions are a relevant health problem. Machine learning (ML) models are increasingly being applied to dermatology as a diagnostic decision support tool using image analysis, especially for skin cancer detection and disease classification. The objective of this study was to perform a prospective validation of an image analysis ML model, which is capable of screening 44 skin diseases, comparing its diagnostic accuracy with that of General Practitioners (GPs) and teledermatology (TD) dermatologists in a real-life setting. Prospective, diagnostic accuracy study including 100 consecutive patients with a skin problem who visited a participating GP in central Catalonia, Spain, between June 2021 and October 2021. The skin issue was first assessed by the GPs. Then an anonymised skin disease picture was taken and uploaded to the ML application, which returned a list with the Top-5 possible diagnosis in order of probability. The same image was then sent to a dermatologist via TD for diagnosis, as per clinical practice. The GPs Top-3, ML model's Top-5 and dermatologist's Top-3 assessments were compared to calculate the accuracy, sensitivity, specificity and diagnostic accuracy of the ML models. The overall Top-1 accuracy of the ML model (39%) was lower than that of GPs (64%) and dermatologists (72%). When the analysis was limited to the diagnoses on which the algorithm had been explicitly trained (n = 82), the balanced Top-1 accuracy of the ML model increased (48%) and in the Top-3 (75%) was comparable to the GPs Top-3 accuracy (76%). The Top-5 accuracy of the ML model (89%) was comparable to the dermatologist Top-3 accuracy (90%). For the different diseases, the sensitivity of the model (Top-3 87% and Top-5 96%) is higher than that of the clinicians (Top-3 GPs 76% and Top-3 dermatologists 84%) only in the benign tumour pathology group, being on the other hand the most prevalent category (n = 53). About the satisfaction of professionals, 92% of the GPs considered it as a useful diagnostic support tool (DST) for the differential diagnosis and in 60% of the cases as an aid in the final diagnosis of the skin lesion. The overall diagnostic accuracy of the model in this study, under real-life conditions, is lower than that of both GPs and dermatologists. This result aligns with the findings of few existing prospective studies conducted under real-life conditions. The outcomes emphasize the significance of involving clinicians in the training of the model and the capability of ML models to assist GPs, particularly in differential diagnosis. Nevertheless, external testing in real-life conditions is crucial for data validation and regulation of these AI diagnostic models before they can be used in primary care.

AI‐powered visual diagnosis of vulvar lichen sclerosus: A pilot study

A Preliminary Study Using High‐Frequency Ultrasound to Evaluate Vulvar Skin With Lichenoid Vulvar Dermatoses

A deep learning-based hybrid artificial intelligence model for the detection and severity assessment of vitiligo lesions

AI-assisted Diagnosis of Vulvovaginal Candidiasis Using Cascaded Neural Networks.

A deep learning algorithm for classification of oral lichen planus lesions from photographic images: A retrospective study

Evaluation of artificial intelligence-powered screening for sexually transmitted infections-related skin lesions using clinical images and metadata

Changes in immunoreactive manganese-superoxide dismutase concentration in human serum after 93 h strenuous physical exercise.

Comparative Analysis of AI Models for Atypical Pigmented Facial Lesion Diagnosis

[Cancer of the kidney: venous staging using magnetic resonance imaging].

Systematic review of deep learning image analyses for the diagnosis and monitoring of skin disease

Single-cell and spatial transcriptomics of vulvar lichen sclerosus reveal multi-compartmental alterations in gene expression and signaling cross-talk

Multi-instance learning based artificial intelligence model to assist vocal fold leukoplakia diagnosis: A multicentre diagnostic study

AI Progress in Skin Lesion Analysis

Human‐multimodal deep learning collaboration in 'precise' diagnosis of lupus erythematosus subtypes and similar skin diseases

Deep Learning Models for Cystoscopic Recognition of Hunner Lesion in Interstitial Cystitis

Exploring the potential of artificial intelligence in improving skin lesion diagnosis in primary care

Artificial Intelligence and Colposcopy: Automatic Identification of Vaginal Squamous Cell Carcinoma Precursors

Diagnostic Performance of Deep Learning Algorithms Applied to Three Common Diagnoses in Dermatopathology

Artificial Intelligence Support for Skin Lesion Triage in Primary Care and Dermatology (Preprint)

Using Artificial Intelligence to Differentiate Mpox from Common Skin Lesions in a Sexual Health Clinic: Development and Evaluation of an Image Recognition Algorithm (Preprint)