Abstract:Objectives Artificial intelligence (AI) has shown promise in improving the performance of fetal ultrasound screening in detecting congenital heart disease (CHD). The effect of giving AI advice to human operators has not been studied in this context. Giving additional information about AI model workings, such as confidence scores for AI predictions, may be a way of improving performance further. Our aims were to investigate whether AI advice improved overall diagnostic accuracy (using a single CHD lesion as an exemplar), and to see what, if any, additional information given to clinicians optimized the overall performance of the clinician‐AI team. Methods An AI model was trained to classify a single fetal CHD lesion (atrioventricular septal defect, AVSD), using a retrospective cohort of 121,130 cardiac four chamber images extracted from 173 ultrasound scan videos (98 with normal hearts, 75 with AVSD). A ResNet50 model architecture was used. Temperature scaling of model prediction probability was performed on a validation set, and gradient‐weighted class activation maps (grad‐CAMs) produced. Ten clinicians (two consultant fetal cardiologists, three trainees in pediatric cardiology, and five fetal cardiac sonographers) were recruited from a center of fetal cardiology to participate. Each participant was shown 2000 fetal four chamber images in a random order (1,000 normal and 1,000 AVSD). The dataset was comprised of 500 images, each shown in four conditions: 1) image alone without AI output; 2) image with binary AI classification; 3) image with AI model confidence; 4) image with gradient‐weighted class activation map image overlays. The clinicians were asked to classify each image as normal or AVSD. Results 20,000 image classifications were recorded from 10 clinicians. The AI model alone achieved an accuracy of 0.798 (95% CI 0.760 – 0.832), sensitivity of 0.868 (95% CI 0.834 – 0.902) and specificity of 0.728 (95% CI 0.702 – 0.754, and the clinicians without AI achieved an accuracy of 0.844 (95% CI 0.834 – 0.854), sensitivity of 0.827 (95% CI 0.795 – 0.858) and specificity of 0.861 (95% CI 0.828 – 0.895). Showing a binary (normal or AVSD) AI model output resulted in significant improvement in accuracy to 0.865 (p <0.001). This effect was seen in both experienced and less experienced participants. Giving incorrect AI advice resulted in significant deterioration in overall accuracy from 0.761 to 0.693 (p <0.001), which was driven by an increase in both type I and type II error by the clinicians. This effect was worsened by showing model confidence (accuracy 0.649, p <0.001) or grad‐CAM (accuracy 0.644, p <0.001). Conclusions AI has the potential to improve performance when used in collaboration with clinicians, even if the model performance does not reach expert level. Giving additional information about model workings such as model confidence and class activation map image overlays did not improve overall performance, and actually worsened performance for images where the AI model was incorrect. This article is protected by copyright. All rights reserved.

Artificial intelligence in fetal echocardiography: Recent advances and future prospects

Advances in the Application of Artificial Intelligence in Fetal Echocardiography

Study on Diagnostic Performance of Fetal Intelligent Navigation Echocardiography for Congenital Heart Defect

Applications of artificial intelligence-powered prenatal diagnosis for congenital heart disease

[Artificial Intelligence Technology in Cardiac Auscultation Screening for Congenital Heart Disease: Present and Future].

Artificial Intelligence in Prenatal Ultrasound Diagnosis

Application Value of Fetal Heart Ultrasound Intelligent Navigation Technique in Display of Key Diagnostic Elements in Rapid Screening Views of Fetal Echocardiography

Novel Foetal Echocardiographic Image Processing Software (5D Heart) Improves the Display of Key Diagnostic Elements in Foetal Echocardiography

Artificial intelligence applications of fetal brain and cardiac MRI

Advancements in Artificial Intelligence for Fetal Neurosonography: A Comprehensive Review

Artificial intelligence in echocardiography: detection, functional evaluation, and disease diagnosis

Review on the intelligent measurement technology of fetal ultrasound image

Application of artificial intelligence in VSD prenatal diagnosis from fetal heart ultrasound images

Artificial Intelligence in Obstetric Anomaly Scan: Heart and Brain

Data for AI in Congenital Heart Defects: Systematic Review

Artificial intelligence applied to fetal MRI: A scoping review of current research

Interaction between clinicians and artificial intelligence to detect fetal atrioventricular septal defects on ultrasound: how can we optimize collaborative performance?

Artificial Intelligence and Echocardiography

Fetal Face: Enhancing 3D Ultrasound Imaging by Postprocessing With AI Applications: Myth, Reality, or Legal Concerns?

The utilization of artificial intelligence in enhancing 3D/4D ultrasound analysis of fetal facial profiles

Artificial Intelligence-Enhanced Echocardiography for Systolic Function Assessment