Confidence Map Research Articles

Artificial intelligence (AI) has shown promise in improving the performance of fetal ultrasound screening in detecting congenital heart disease (CHD). The effect of giving AI advice to human operators has not been studied in this context. Giving additional information about AI model workings, such as confidence scores for AI predictions, may be a way of further improving performance. Our aims were to investigate whether AI advice improved overall diagnostic accuracy (using a single CHD lesion as an exemplar), and to determine what, if any, additional information given to clinicians optimized the overall performance of the clinician-AI team. An AI model was trained to classify a single fetal CHD lesion (atrioventricular septal defect (AVSD)), using a retrospective cohort of 121 130 cardiac four-chamber images extracted from 173 ultrasound scan videos (98 with normal hearts, 75 with AVSD); a ResNet50 model architecture was used. Temperature scaling of model prediction probability was performed on a validation set, and gradient-weighted class activation maps (grad-CAMs) produced. Ten clinicians (two consultant fetal cardiologists, three trainees in pediatric cardiology and five fetal cardiac sonographers) were recruited from a center of fetal cardiology to participate. Each participant was shown 2000 fetal four-chamber images in a random order (1000 normal and 1000 AVSD). The dataset comprised 500 images, each shown in four conditions: (1) image alone without AI output; (2) image with binary AI classification; (3) image with AI model confidence; and (4) image with grad-CAM image overlays. The clinicians were asked to classify each image as normal or AVSD. A total of 20 000 image classifications were recorded from 10 clinicians. The AI model alone achieved an accuracy of 0.798 (95% CI, 0.760-0.832), a sensitivity of 0.868 (95% CI, 0.834-0.902) and a specificity of 0.728 (95% CI, 0.702-0.754), and the clinicians without AI achieved an accuracy of 0.844 (95% CI, 0.834-0.854), a sensitivity of 0.827 (95% CI, 0.795-0.858) and a specificity of 0.861 (95% CI, 0.828-0.895). Showing a binary (normal or AVSD) AI model output resulted in significant improvement in accuracy to 0.865 (P < 0.001). This effect was seen in both experienced and less-experienced participants. Giving incorrect AI advice resulted in a significant deterioration in overall accuracy, from 0.761 to 0.693 (P < 0.001), which was driven by an increase in both Type-I and Type-II errors by the clinicians. This effect was worsened by showing model confidence (accuracy, 0.649; P < 0.001) or grad-CAM (accuracy, 0.644; P < 0.001). AI has the potential to improve performance when used in collaboration with clinicians, even if the model performance does not reach expert level. Giving additional information about model workings such as model confidence and class activation map image overlays did not improve overall performance, and actually worsened performance for images for which the AI model was incorrect. © 2024 The Authors. Ultrasound in Obstetrics & Gynecology published by John Wiley & Sons Ltd on behalf of International Society of Ultrasound in Obstetrics and Gynecology.

Read full abstract

Various subjective and objective methods have been proposed to identify which interictal epileptiform discharge (IED)-related EEG-fMRI results are more likely to delineate seizure generating tissue in patients with drug-resistant focal epilepsy for the purposes of surgical planning. In this intracranial EEG-fMRI study, we evaluated the utility of these methods to localize clinically relevant regions pre-operatively and compared the extent of resection of these areas to post-operative outcome. Seventy patients admitted for intracranial video-EEG monitoring were recruited for a simultaneous intracranial EEG-fMRI study. For all analyses of blood oxygen level-dependent responses associated with IEDs, an experienced epileptologist identified the most Clinically Relevant brain activation cluster using available clinical information. The Maximum cluster (the cluster with the highest z-score) was also identified for all analyses and assigned to one of three confidence levels (low, medium, or high) based on the difference of the peak z-scores between the Maximum and Second Maximum cluster (the cluster with the second highest peak z-value). The distance was measured and compared between the peak voxel of the aforementioned clusters and the electrode contacts where the interictal discharge and seizure onset were recorded. In patients who subsequently underwent epilepsy surgery, the spatial concordance between the aforementioned clusters and the area of resection was determined and compared to post-operative outcome. We evaluated 106 different IEDs in 70 patients. Both subjective (identification of the Clinically Relevant cluster) and objective (Maximum cluster much more significant than the second maximum cluster) methods of culling non-localizing EEG-fMRI activation maps increased the spatial concordance between these clusters and the corresponding IED or seizure onset zone contacts. However, only the objective methods of identifying medium and high confidence maps resulted in a significant association between resection of the peak voxel of the Maximum cluster and post-operative outcome. Resection of this area was associated with good post-operative outcomes but was not sufficient for seizure freedom. On the other hand, we found that failure to resect the medium and high confidence Maximum clusters was associated with a poor post-surgical outcome (negative predictive value = 1.0, sensitivity = 1.0). Objective methods to identify higher confidence EEG-fMRI results are needed to localize areas necessary for good post-operative outcomes. However, resection of the peak voxel within higher confidence Maximum clusters is not sufficient for good outcomes. Conversely, failure to resect the peak voxel in these clusters is associated with a poor post-surgical outcome.

Read full abstract

Confidence Map Research Articles

Related Topics

Articles published on Confidence Map

Incorporating modelling uncertainty and prior knowledge into landslide susceptibility mapping using Bayesian neural networks

Do as Sonographers Think: Contrast-Enhanced Ultrasound for Thyroid Nodules Diagnosis via Microvascular Infiltrative Awareness.

Quantifying uncertainty in landslide susceptibility mapping due to sampling randomness

AI-smartphone markerless motion capturing of hip, knee, and ankle joint kinematics during countermovement jumps.

Feature-Level Fusion Multi-Sensor Aggregation Temporal Network for Smartphone-Based Human Activity Recognition

DVDS: A deep visual dynamic slam system

Automated confidence estimation in deep learning auto-segmentation for brain organs at risk on MRI for radiotherapy.

Efficient occlusion avoidance based on active deep sensing for harvesting robots

Utilizing RT-DETR Model for Fruit Calorie Estimation from Digital Images

Multi-Task Credible Pseudo-Label Learning for Semi-Supervised Crowd Counting.

Identifying and fitting eclipse maps of exoplanets with cross-validation

Interaction between clinicians and artificial intelligence to detect fetal atrioventricular septal defects on ultrasound: how can we optimize collaborative performance?

Mapping interictal discharges using intracranial EEG-fMRI to predict postsurgical outcomes.

Where is my attention? An explainable AI exploration in water detection from SAR imagery

Data-limited and imbalanced bladder wall segmentation with confidence map-guided residual networks via transfer learning

Confidence maps for reliable estimation of proton density fat fraction and in the liver.

Attentional pixel-wise deformation for pose-based human image generation

Depth Completion with Multiple Balanced Bases and Confidence for Dense Monocular SLAM.

A Deformation Field-Based Approach for Expression Processing in Head Motion Correction

Preliminary results of automatic cotton crops mapping using remote sensing data

Lead the way for us

Editage

Paperpal

R Discovery

Mind the Graph

Confidence Map Research Articles

Related Topics

Articles published on Confidence Map

Incorporating modelling uncertainty and prior knowledge into landslide susceptibility mapping using Bayesian neural networks

Do as Sonographers Think: Contrast-Enhanced Ultrasound for Thyroid Nodules Diagnosis via Microvascular Infiltrative Awareness.

Quantifying uncertainty in landslide susceptibility mapping due to sampling randomness

AI-smartphone markerless motion capturing of hip, knee, and ankle joint kinematics during countermovement jumps.

Feature-Level Fusion Multi-Sensor Aggregation Temporal Network for Smartphone-Based Human Activity Recognition

DVDS: A deep visual dynamic slam system

Automated confidence estimation in deep learning auto-segmentation for brain organs at risk on MRI for radiotherapy.

Efficient occlusion avoidance based on active deep sensing for harvesting robots

Utilizing RT-DETR Model for Fruit Calorie Estimation from Digital Images

Multi-Task Credible Pseudo-Label Learning for Semi-Supervised Crowd Counting.

Identifying and fitting eclipse maps of exoplanets with cross-validation

Interaction between clinicians and artificial intelligence to detect fetal atrioventricular septal defects on ultrasound: how can we optimize collaborative performance?

Mapping interictal discharges using intracranial EEG-fMRI to predict postsurgical outcomes.

Where is my attention? An explainable AI exploration in water detection from SAR imagery

Data-limited and imbalanced bladder wall segmentation with confidence map-guided residual networks via transfer learning

Confidence maps for reliable estimation of proton density fat fraction and in the liver.

Attentional pixel-wise deformation for pose-based human image generation

Depth Completion with Multiple Balanced Bases and Confidence for Dense Monocular SLAM.

A Deformation Field-Based Approach for Expression Processing in Head Motion Correction

Preliminary results of automatic cotton crops mapping using remote sensing data