Review

. 2023 Nov 5:15:100348.

doi: 10.1016/j.jpi.2023.100348. eCollection 2024 Dec.

Performance of externally validated machine learning models based on histopathology images for the diagnosis, classification, prognosis, or treatment outcome prediction in female breast cancer: A systematic review

Ricardo Gonzalez^{1

2}, Peyman Nejat³, Ashirbani Saha^{4

5}, Clinton J V Campbell^{6

7}, Andrew P Norgan⁸, Cynthia Lokker⁹

Affiliations

¹ DeGroote School of Business, McMaster University, Hamilton, Ontario, Canada.
² Division of Computational Pathology and Artificial Intelligence, Department of Laboratory Medicine and Pathology, Mayo Clinic, Rochester, MN, United States.
³ Department of Artificial Intelligence and Informatics, Mayo Clinic, Rochester, MN, United States.
⁴ Department of Oncology, Faculty of Health Sciences, McMaster University, Hamilton, Ontario, Canada.
⁵ Escarpment Cancer Research Institute, McMaster University and Hamilton Health Sciences, Hamilton, Ontario, Canada.
⁶ William Osler Health System, Brampton, Ontario, Canada.
⁷ Department of Pathology and Molecular Medicine, Faculty of Health Sciences, McMaster University, Hamilton, Ontario, Canada.
⁸ Department of Laboratory Medicine and Pathology, Mayo Clinic, Rochester, MN, United States.
⁹ Health Information Research Unit, Department of Health Research Methods, Evidence and Impact, McMaster University, Hamilton, Ontario, Canada.

PMID: 38089005
PMCID: PMC10714242
DOI: 10.1016/j.jpi.2023.100348

Review

Performance of externally validated machine learning models based on histopathology images for the diagnosis, classification, prognosis, or treatment outcome prediction in female breast cancer: A systematic review

Ricardo Gonzalez et al. J Pathol Inform. 2023.

. 2023 Nov 5:15:100348.

doi: 10.1016/j.jpi.2023.100348. eCollection 2024 Dec.

Authors

Ricardo Gonzalez^{1

2}, Peyman Nejat³, Ashirbani Saha^{4

5}, Clinton J V Campbell^{6

7}, Andrew P Norgan⁸, Cynthia Lokker⁹

Affiliations

¹ DeGroote School of Business, McMaster University, Hamilton, Ontario, Canada.
² Division of Computational Pathology and Artificial Intelligence, Department of Laboratory Medicine and Pathology, Mayo Clinic, Rochester, MN, United States.
³ Department of Artificial Intelligence and Informatics, Mayo Clinic, Rochester, MN, United States.
⁴ Department of Oncology, Faculty of Health Sciences, McMaster University, Hamilton, Ontario, Canada.
⁵ Escarpment Cancer Research Institute, McMaster University and Hamilton Health Sciences, Hamilton, Ontario, Canada.
⁶ William Osler Health System, Brampton, Ontario, Canada.
⁷ Department of Pathology and Molecular Medicine, Faculty of Health Sciences, McMaster University, Hamilton, Ontario, Canada.
⁸ Department of Laboratory Medicine and Pathology, Mayo Clinic, Rochester, MN, United States.
⁹ Health Information Research Unit, Department of Health Research Methods, Evidence and Impact, McMaster University, Hamilton, Ontario, Canada.

PMID: 38089005
PMCID: PMC10714242
DOI: 10.1016/j.jpi.2023.100348

Abstract

Numerous machine learning (ML) models have been developed for breast cancer using various types of data. Successful external validation (EV) of ML models is important evidence of their generalizability. The aim of this systematic review was to assess the performance of externally validated ML models based on histopathology images for diagnosis, classification, prognosis, or treatment outcome prediction in female breast cancer. A systematic search of MEDLINE, EMBASE, CINAHL, IEEE, MICCAI, and SPIE conferences was performed for studies published between January 2010 and February 2022. The Prediction Model Risk of Bias Assessment Tool (PROBAST) was employed, and the results were narratively described. Of the 2011 non-duplicated citations, 8 journal articles and 2 conference proceedings met inclusion criteria. Three studies externally validated ML models for diagnosis, 4 for classification, 2 for prognosis, and 1 for both classification and prognosis. Most studies used Convolutional Neural Networks and one used logistic regression algorithms. For diagnostic/classification models, the most common performance metrics reported in the EV were accuracy and area under the curve, which were greater than 87% and 90%, respectively, using pathologists' annotations/diagnoses as ground truth. The hazard ratios in the EV of prognostic ML models were between 1.7 (95% CI, 1.2-2.6) and 1.8 (95% CI, 1.3-2.7) to predict distant disease-free survival; 1.91 (95% CI, 1.11-3.29) for recurrence, and between 0.09 (95% CI, 0.01-0.70) and 0.65 (95% CI, 0.43-0.98) for overall survival, using clinical data as ground truth. Despite EV being an important step before the clinical application of a ML model, it hasn't been performed routinely. The large variability in the training/validation datasets, methods, performance metrics, and reported information limited the comparison of the models and the analysis of their results. Increasing the availability of validation datasets and implementing standardized methods and reporting protocols may facilitate future analyses.

Keywords: Breast neoplasms; Machine learning; Pathology; Systematic review; Validation studies.

PubMed Disclaimer

Conflict of interest statement

Authors do not have any competing interests to declare. This research did not receive any specific grant from funding agencies in the public, commercial, or not-for-profit sectors.

Figures

**Fig. 1**
PRISMA flow diagram of the studies identification process for the systematic review.

**Fig. 2**
Distribution of the records reviewed during the Title/Abstract screening by year (January 1, 2010–February 28, 2022).

**Fig. 3**
Prediction model Risk Of Bias Assessment Tool (PROBAST) Graphical presentation—(1) risk of bias results and (2) applicability.

**Fig. 4**
ML models site-specific iterative development steps review.

**Fig. 5**
Record that passed the Title/Abstract screening and the full-text screening.

**Fig. 6**
Record that passed the Title/Abstract screening and the full-text screening.

**Fig. 7**
Record that passed the Title/Abstract screening but did not pass the full-text screening (because EV was not performed).

**Fig. 8**
Record that passed the Title/Abstract screening but did not pass the full-text screening (because EV was not performed).

**Fig. 9**
Record that did not pass the Title/Abstract screening (because EV was not performed).

**Fig. 10**
Record that did not pass the Title/Abstract screening (because the model was only trained with microarray and clinical data).

**Fig. 11**
Record that did not pass the Title/Abstract screening (because the model was only trained with ultrasound images).

**Fig. 12**
Record that did not pass the Title/Abstract screening (because the model was developed to detect mitoses).

**Fig. 13**
Record that did not pass the Title/Abstract screening (because the model was developed to predict the expression of biomarkers).

See this image and copyright information in PMC

References

1. Sung H., Ferlay J., Siegel R.L., et al. Global cancer statistics 2020: GLOBOCAN estimates of incidence and mortality worldwide for 36 cancers in 185 countries. CA Cancer J Clin. 2021;71(3):209–249. doi: 10.3322/caac.21660. - DOI - PubMed
1. Khened M., Kori A., Rajkumar H., Krishnamurthi G., Srinivasan B. A generalized deep learning framework for whole-slide image segmentation and analysis. Sci Rep. 2021;11(1):11579. doi: 10.1038/s41598-021-90444-8. - DOI - PMC - PubMed
1. Leong A.S.Y., Zhuang Z. The changing role of pathology in breast cancer diagnosis and treatment. Pathobiol J Immunopathol Mol Cell Biol. 2011;78(2):99–114. doi: 10.1159/000292644. - DOI - PMC - PubMed
1. Nam S., Chong Y., Jung C.K., et al. Introduction to digital pathology and computer-aided pathology. J Pathol Transl Med. 2020;54(2):125–134. doi: 10.4132/jptm.2019.12.31. - DOI - PMC - PubMed
1. Randell R., Ruddle R.A., Treanor D. Barriers and facilitators to the introduction of digital pathology for diagnostic work. Stud Health Technol Inform. 2015;216:443–447. - PubMed

Publication types

Actions

LinkOut - more resources

Full Text Sources
Miscellaneous
- NCI CPTAC Assay Portal

Save citation to file

Email citation

Add to Collections

Add to My Bibliography

Your saved search

Create a file for external citation management software

Your RSS Feed

Performance of externally validated machine learning models based on histopathology images for the diagnosis, classification, prognosis, or treatment outcome prediction in female breast cancer: A systematic review

Affiliations

Performance of externally validated machine learning models based on histopathology images for the diagnosis, classification, prognosis, or treatment outcome prediction in female breast cancer: A systematic review

Authors

Affiliations

Abstract

Conflict of interest statement

Figures

References

Publication types

LinkOut - more resources

Full Text Sources

Miscellaneous