. 2022 Jun;4(6):e406-e414.

doi: 10.1016/S2589-7500(22)00063-2. Epub 2022 May 11.

AI recognition of patient race in medical imaging: a modelling study

Judy Wawira Gichoya¹, Imon Banerjee², Ananth Reddy Bhimireddy³, John L Burns⁴, Leo Anthony Celi⁵, Li-Ching Chen⁶, Ramon Correa², Natalie Dullerud⁷, Marzyeh Ghassemi⁸, Shih-Cheng Huang⁹, Po-Chih Kuo⁶, Matthew P Lungren⁹, Lyle J Palmer¹⁰, Brandon J Price¹¹, Saptarshi Purkayastha⁴, Ayis T Pyrros¹², Lauren Oakden-Rayner¹³, Chima Okechukwu¹⁴, Laleh Seyyed-Kalantari¹⁵, Hari Trivedi³, Ryan Wang⁶, Zachary Zaiman¹⁶, Haoran Zhang⁷

Affiliations

¹ Department of Radiology, Emory University, Atlanta, GA, USA. Electronic address: judywawira@emory.edu.
² School of Computing, Informatics, and Decision Systems Engineering, Arizona State University, Tempe, AZ, USA.
³ Department of Radiology, Emory University, Atlanta, GA, USA.
⁴ School of Informatics and Computing, Indiana University-Purdue University, Indianapolis, IN, USA.
⁵ Institute for Medical Engineering and Science, Massachusetts Institute of Technology, Cambridge, MA, USA; Department of Medicine, Beth Israel Deaconess Medical Center, Boston, MA, USA.
⁶ Department of Computer Science, National Tsing Hua University, Hsinchu, Taiwan.
⁷ Department of Computer Science, University of Toronto, Toronto, ON, Canada.
⁸ Institute for Medical Engineering and Science, Massachusetts Institute of Technology, Cambridge, MA, USA; Department of Electrical Engineering and Computer Science, Massachusetts Institute of Technology, Cambridge, MA, USA.
⁹ Stanford University School of Medicine, Palo Alto, CA, USA.
¹⁰ Australian Institute for Machine Learning, University of Adelaide, Adelaide, SA, Australia; School of Public Health, University of Adelaide, Adelaide, SA, Australia.
¹¹ Florida State University College of Medicine, Tallahassee, FL, USA.
¹² Dupage Medical Group, Hinsdale, IL, USA.
¹³ Australian Institute for Machine Learning, University of Adelaide, Adelaide, SA, Australia.
¹⁴ Department of Computer Science, Georgia Institute of Technology, Atlanta, GA, USA.
¹⁵ Department of Computer Science, University of Toronto, Toronto, ON, Canada; Lunenfeld-Tanenbaum Research Institute, Sinai Health, Toronto, ON, Canada; Vector Institute for Artificial Intelligence, Toronto, ON, Canada.
¹⁶ Department of Computer Science, Emory University, Atlanta, GA, USA.

PMID: 35568690
PMCID: PMC9650160
DOI: 10.1016/S2589-7500(22)00063-2

AI recognition of patient race in medical imaging: a modelling study

Judy Wawira Gichoya et al. Lancet Digit Health. 2022 Jun.

. 2022 Jun;4(6):e406-e414.

doi: 10.1016/S2589-7500(22)00063-2. Epub 2022 May 11.

Authors

Affiliations

¹ Department of Radiology, Emory University, Atlanta, GA, USA. Electronic address: judywawira@emory.edu.
² School of Computing, Informatics, and Decision Systems Engineering, Arizona State University, Tempe, AZ, USA.
³ Department of Radiology, Emory University, Atlanta, GA, USA.
⁴ School of Informatics and Computing, Indiana University-Purdue University, Indianapolis, IN, USA.
⁵ Institute for Medical Engineering and Science, Massachusetts Institute of Technology, Cambridge, MA, USA; Department of Medicine, Beth Israel Deaconess Medical Center, Boston, MA, USA.
⁶ Department of Computer Science, National Tsing Hua University, Hsinchu, Taiwan.
⁷ Department of Computer Science, University of Toronto, Toronto, ON, Canada.
⁸ Institute for Medical Engineering and Science, Massachusetts Institute of Technology, Cambridge, MA, USA; Department of Electrical Engineering and Computer Science, Massachusetts Institute of Technology, Cambridge, MA, USA.
⁹ Stanford University School of Medicine, Palo Alto, CA, USA.
¹⁰ Australian Institute for Machine Learning, University of Adelaide, Adelaide, SA, Australia; School of Public Health, University of Adelaide, Adelaide, SA, Australia.
¹¹ Florida State University College of Medicine, Tallahassee, FL, USA.
¹² Dupage Medical Group, Hinsdale, IL, USA.
¹³ Australian Institute for Machine Learning, University of Adelaide, Adelaide, SA, Australia.
¹⁴ Department of Computer Science, Georgia Institute of Technology, Atlanta, GA, USA.
¹⁵ Department of Computer Science, University of Toronto, Toronto, ON, Canada; Lunenfeld-Tanenbaum Research Institute, Sinai Health, Toronto, ON, Canada; Vector Institute for Artificial Intelligence, Toronto, ON, Canada.
¹⁶ Department of Computer Science, Emory University, Atlanta, GA, USA.

PMID: 35568690
PMCID: PMC9650160
DOI: 10.1016/S2589-7500(22)00063-2

Abstract

Background: Previous studies in medical imaging have shown disparate abilities of artificial intelligence (AI) to detect a person's race, yet there is no known correlation for race on medical imaging that would be obvious to human experts when interpreting the images. We aimed to conduct a comprehensive evaluation of the ability of AI to recognise a patient's racial identity from medical images.

Methods: Using private (Emory CXR, Emory Chest CT, Emory Cervical Spine, and Emory Mammogram) and public (MIMIC-CXR, CheXpert, National Lung Cancer Screening Trial, RSNA Pulmonary Embolism CT, and Digital Hand Atlas) datasets, we evaluated, first, performance quantification of deep learning models in detecting race from medical images, including the ability of these models to generalise to external environments and across multiple imaging modalities. Second, we assessed possible confounding of anatomic and phenotypic population features by assessing the ability of these hypothesised confounders to detect race in isolation using regression models, and by re-evaluating the deep learning models by testing them on datasets stratified by these hypothesised confounding variables. Last, by exploring the effect of image corruptions on model performance, we investigated the underlying mechanism by which AI models can recognise race.

Findings: In our study, we show that standard AI deep learning models can be trained to predict race from medical images with high performance across multiple imaging modalities, which was sustained under external validation conditions (x-ray imaging [area under the receiver operating characteristics curve (AUC) range 0·91-0·99], CT chest imaging [0·87-0·96], and mammography [0·81]). We also showed that this detection is not due to proxies or imaging-related surrogate covariates for race (eg, performance of possible confounders: body-mass index [AUC 0·55], disease distribution [0·61], and breast density [0·61]). Finally, we provide evidence to show that the ability of AI deep learning models persisted over all anatomical regions and frequency spectrums of the images, suggesting the efforts to control this behaviour when it is undesirable will be challenging and demand further study.

Interpretation: The results from our study emphasise that the ability of AI deep learning models to predict self-reported race is itself not the issue of importance. However, our finding that AI can accurately predict self-reported race, even from corrupted, cropped, and noised medical images, often when clinical experts cannot, creates an enormous risk for all model deployments in medical imaging.

Funding: National Institute of Biomedical Imaging and Bioengineering, MIDRC grant of National Institutes of Health, US National Science Foundation, National Library of Medicine of the National Institutes of Health, and Taiwan Ministry of Science and Technology.

PubMed Disclaimer

Conflict of interest statement

Declaration of interests MG has received speaker fees for a Harvard Medical School executive education class. HT has received consulting fees from Sirona medical, Arterys, and Biodata consortium. HT also owns lightbox AI, which provides expert annotation of medical images for radiology AI. MPL has received consulting fees from Bayer, Microsoft, Phillips, and Nines. MPL also owns stocks in Nines, SegMed, and Centaur. LAC has received support to attend meetings from MISTI Global Seed Funds. ATP has received payment for expert testimony from NCMIC insurance company. ATP also has a pending institutional patent for comorbidity prediction from radiology images. All other authors declare no competing interests.

Figures

**Figure 1:. The effect on model performance of adding a low-pass filter and a high-pass filter for various diameters in the MXR dataset**
MXR=MIMIC-CXR dataset.

**Figure 2:. Samples of the images after low-pass filters and high-pass filters in MXR dataset**
HPF=high-pass filtering. LPF=low-pass filtering. MXR=MIMIC-CXR dataset.

See this image and copyright information in PMC

Comment in

AI models in health care are not colour blind and we should not be either.
Wiens J, Creary M, Sjoding MW. Wiens J, et al. Lancet Digit Health. 2022 Jun;4(6):e399-e400. doi: 10.1016/S2589-7500(22)00092-9. Epub 2022 May 11. Lancet Digit Health. 2022. PMID: 35568691 No abstract available.
Racial Identity Remains Embedded within Medical Imaging Data.
Singh P, Savjani R. Singh P, et al. Radiol Imaging Cancer. 2022 Jul;4(4):e229014. doi: 10.1148/rycan.229014. Radiol Imaging Cancer. 2022. PMID: 35866890 Free PMC article. No abstract available.
Beyond the AJR: Robust Ability of Artificial Intelligence to Detect Race Underscores the Need for Inclusivity and Transparency.
Shah C, Chen PH. Shah C, et al. AJR Am J Roentgenol. 2023 Mar;220(3):449. doi: 10.2214/AJR.22.28293. Epub 2022 Jul 27. AJR Am J Roentgenol. 2023. PMID: 35895299 No abstract available.

References

1. Bender EM, Gebru T, McMillan-Major A, Shmitchell S. On the dangers of stochastic parrots: can language models be too big? FAccT ‘21; March 3–10, 2021. 10.1145/3442188.3445922. - DOI
1. Angwin J, Larson J, Mattu S, Kirchner L. Machine bias. https://www.propublica.org/article/machine-bias-risk-assessments-in-crim... (accessed april 25, 2022).
1. Koenecke A, Nam A, Lake E, et al. Racial disparities in automated speech recognition. Proc Natl Acad Sci USA 2020; 117: 7684–89. - PMC - PubMed
1. Buolamwini J, Gebru T. Gender shades: intersectional accuracy disparities in commercial gender classification. PMLR 2018; 81: 77–91.
1. Adamson AS, Smith A. Machine learning and health care disparities in dermatology. JAMA Dermatol 2018; 154: 1247–48. - PubMed

Publication types

Actions
Actions
Actions

MeSH terms

Actions
Actions
Actions
Actions
Actions
Actions

Grants and funding

LinkOut - more resources

Full Text Sources
Medical
- MedlinePlus Health Information

Save citation to file

Email citation

Add to Collections

Add to My Bibliography

Your saved search

Create a file for external citation management software

Your RSS Feed

AI recognition of patient race in medical imaging: a modelling study

Affiliations

AI recognition of patient race in medical imaging: a modelling study

Authors

Affiliations

Abstract

Conflict of interest statement

Figures

Comment in

References

Publication types

MeSH terms

Grants and funding

LinkOut - more resources

Full Text Sources

Medical