CORTEXA
← Browse
openalexDiagnostics2026-07-23Cited by 0

Artificial Intelligence for Breast MRI Lesion Classification: A Targeted Evidence Synthesis and Meta-Analysis of Discriminative Performance and Heterogeneity

Romuald Ferré, Thad Benefield, Cherie M. Kuzmiak

Background/Objectives: The paper aimed to synthesize the diagnostic performance of artificial intelligence (AI) methods for classifying breast lesions on contrast-enhanced breast MRI and to estimate a pooled area under the receiver operating characteristic curve (AUC). Methods: This targeted evidence synthesis and meta-analysis was informed by PRISMA 2020 reporting principles where applicable. Eligible studies were drawn from an investigator-supplied corpus of 12 primary manuscripts and assessed against predefined criteria. MEDLINE/PubMed, Embase, and Web of Science were consulted through October 2025 to contextualize the literature and verify bibliographic and study details; additional database records were not screened for eligibility. We included studies applying machine learning or deep learning to contrast-enhanced breast MRI for benign-versus-malignant lesion classification and reporting an AUC on an independent test set, external validation set, or patient-wise cross-validation. AUCs were pooled on the logit scale using an inverse-variance DerSimonian–Laird random-effects model, and heterogeneity was quantified using I2. Results: Nine studies met the criteria for quantitative synthesis (evaluation-set sizes, 60–3936). The pooled random-effects AUC was 0.898 (95% CI, 0.875–0.918), with substantial heterogeneity (I2 = 88.1%) and a 95% prediction interval of 0.824–0.943, indicating that performance may vary meaningfully across settings. Conclusions: AI models showed promising discriminative performance within this targeted corpus, but substantial heterogeneity, differences in unit of analysis, approximated variance estimates, and limited external institutional validation temper confidence in generalizability. The pooled AUC should be interpreted descriptively, and future studies should prioritize rigorous multi-institutional external validation, transparent reporting, and prospective reader- or workflow-impact evaluation before routine clinical deployment.

View free PDFSource page

Related papers

crossrefDiagnostics2025-04-09Cited by 10

Comparative Evaluation of Machine Learning-Based Radiomics and Deep Learning for Breast Lesion Classification in Mammography

Alessandro Stefano, Fabiano Bini, Eleonora Giovagnoli, Mariangela Dimarco, Nicolò Lauciello, Daniela Narbonese, et al.

Background: Breast cancer is the second leading cause of cancer-related mortality among women, accounting for 12% of cases. Early diagnosis, based on the identification of radiological features, such as masses and microcalcifications in mammograms, is crucial for reducing mortali…

View free PDFSource page
crossrefDiagnostics2026-06-18

Artificial Intelligence, Deep Learning, and Computer Vision in Hysteroscopy: A Systematic Review

Rafał Watrowski, Attilio Di Spiezio Sardo, Peter Török, Andrea Rosati, Stoyan Kostov, Ibrahim Alkatout, et al.

Background/Objectives: Hysteroscopy is the gold standard for visualization and treatment of intrauterine pathology. Because hysteroscopic interpretation remains operator-dependent, artificial intelligence (AI) has been evaluated as a tool to improve consistency, lesion recognitio…

View free PDFSource page
crossrefDiagnostics2026-06-11

Artificial Intelligence in Orofacial Pain: Diagnostic and Predictive Performance Across Machine Learning and Deep Learning Models

Laura Iosif, Marina Imre, Andreea Gabriela Wagner, Ana Maria Cristina Țâncu, Andreea Cristiana Didilescu, Hendrik Simon Brand, et al.

Orofacial pain (OFP) includes a broad spectrum of odontogenic and non-odontogenic conditions with overlapping clinical features that often limit diagnostic accuracy, driving increasing interest in artificial intelligence (AI) as a tool to enhance diagnostic precision and support…

View free PDFSource page
crossrefDiagnostics2026-04-01

Generative Artificial Intelligence vs. Transformer and Benchmarking Against Deep/Machine Learning: Classification and Scientific Validation of Heart Failure Patients Using Women’s Transcriptomic Gene Data

Ekta Tiwari, Dipti Shrimankar, Krish Chaudhary, Luca Saba, Jasjit S. Suri

Backgrounds: Accurate early classification of heart failure (HF) in women is challenging due to sex-specific gene expression and disease patterns, which traditional models overlook. We propose a generative artificial intelligence (GenAI)-based model for the classification of HF p…

View free PDFSource page
crossrefDiagnostics2022-01-13Cited by 6

Applied Machine Learning in Spiral Breast-CT: Can We Train a Deep Convolutional Neural Network for Automatic, Standardized and Observer Independent Classification of Breast Density?

Anna Landsmann, Jann Wieler, Patryk Hejduk, Alexander Ciritsis, Karol Borkowski, Cristina Rossi, et al.

The aim of this study was to investigate the potential of a machine learning algorithm to accurately classify parenchymal density in spiral breast-CT (BCT), using a deep convolutional neural network (dCNN). In this retrospectively designed study, 634 examinations of 317 patients…

View free PDFSource page
crossrefDiagnostics2024-01-12Cited by 54

From Machine Learning to Patient Outcomes: A Comprehensive Review of AI in Pancreatic Cancer

Satvik Tripathi, Azadeh Tabari, Arian Mansur, Harika Dabbara, Christopher P. Bridge, Dania Daye

Pancreatic cancer is a highly aggressive and difficult-to-detect cancer with a poor prognosis. Late diagnosis is common due to a lack of early symptoms, specific markers, and the challenging location of the pancreas. Imaging technologies have improved diagnosis, but there is stil…

View free PDFSource page