A Multimodality Machine Learning Approach to Differentiate Severe and Nonsevere COVID-19: Model Development and Validation
- PMID: 33714935
- PMCID: PMC8030658
- DOI: 10.2196/23948
A Multimodality Machine Learning Approach to Differentiate Severe and Nonsevere COVID-19: Model Development and Validation
Abstract
Background: Effectively and efficiently diagnosing patients who have COVID-19 with the accurate clinical type of the disease is essential to achieve optimal outcomes for the patients as well as to reduce the risk of overloading the health care system. Currently, severe and nonsevere COVID-19 types are differentiated by only a few features, which do not comprehensively characterize the complicated pathological, physiological, and immunological responses to SARS-CoV-2 infection in the different disease types. In addition, these type-defining features may not be readily testable at the time of diagnosis.
Objective: In this study, we aimed to use a machine learning approach to understand COVID-19 more comprehensively, accurately differentiate severe and nonsevere COVID-19 clinical types based on multiple medical features, and provide reliable predictions of the clinical type of the disease.
Methods: For this study, we recruited 214 confirmed patients with nonsevere COVID-19 and 148 patients with severe COVID-19. The clinical characteristics (26 features) and laboratory test results (26 features) upon admission were acquired as two input modalities. Exploratory analyses demonstrated that these features differed substantially between two clinical types. Machine learning random forest models based on all the features in each modality as well as on the top 5 features in each modality combined were developed and validated to differentiate COVID-19 clinical types.
Results: Using clinical and laboratory results independently as input, the random forest models achieved >90% and >95% predictive accuracy, respectively. The importance scores of the input features were further evaluated, and the top 5 features from each modality were identified (age, hypertension, cardiovascular disease, gender, and diabetes for the clinical features modality, and dimerized plasmin fragment D, high sensitivity troponin I, absolute neutrophil count, interleukin 6, and lactate dehydrogenase for the laboratory testing modality, in descending order). Using these top 10 multimodal features as the only input instead of all 52 features combined, the random forest model was able to achieve 97% predictive accuracy.
Conclusions: Our findings shed light on how the human body reacts to SARS-CoV-2 infection as a unit and provide insights on effectively evaluating the disease severity of patients with COVID-19 based on more common medical features when gold standard features are not available. We suggest that clinical information can be used as an initial screening tool for self-evaluation and triage, while laboratory test results should be applied when accuracy is the priority.
Keywords: COVID-19; classification; clinical type; decision support; diagnosis; machine learning; multimodality; prediction; reliable.
©Yuanfang Chen, Liu Ouyang, Forrest S Bao, Qian Li, Lei Han, Hengdong Zhang, Baoli Zhu, Yaorong Ge, Patrick Robinson, Ming Xu, Jie Liu, Shi Chen. Originally published in the Journal of Medical Internet Research (http://www.jmir.org), 07.04.2021.
Conflict of interest statement
Conflicts of Interest: None declared.
Figures




Similar articles
-
Accurately Differentiating Between Patients With COVID-19, Patients With Other Viral Infections, and Healthy Individuals: Multimodal Late Fusion Learning Approach.J Med Internet Res. 2021 Jan 6;23(1):e25535. doi: 10.2196/25535. J Med Internet Res. 2021. PMID: 33404516 Free PMC article.
-
Clinical and inflammatory features based machine learning model for fatal risk prediction of hospitalized COVID-19 patients: results from a retrospective cohort study.Ann Med. 2021 Dec;53(1):257-266. doi: 10.1080/07853890.2020.1868564. Ann Med. 2021. PMID: 33410720 Free PMC article.
-
Machine Learning Applied to Clinical Laboratory Data in Spain for COVID-19 Outcome Prediction: Model Development and Validation.J Med Internet Res. 2021 Apr 14;23(4):e26211. doi: 10.2196/26211. J Med Internet Res. 2021. PMID: 33793407 Free PMC article.
-
Benchmarking of Machine Learning classifiers on plasma proteomic for COVID-19 severity prediction through interpretable artificial intelligence.Artif Intell Med. 2023 Mar;137:102490. doi: 10.1016/j.artmed.2023.102490. Epub 2023 Jan 18. Artif Intell Med. 2023. PMID: 36868685 Free PMC article. Review.
-
Potential solutions for screening, triage, and severity scoring of suspected COVID-19 positive patients in low-resource settings: a scoping review.BMJ Open. 2021 Sep 15;11(9):e046130. doi: 10.1136/bmjopen-2020-046130. BMJ Open. 2021. PMID: 34526332 Free PMC article.
Cited by
-
Derivation and external validation of a nomogram predicting the occurrence of severe illness among hospitalized coronavirus disease 2019 patients: a 2020 Chinese multicenter retrospective study.J Thorac Dis. 2023 Dec 30;15(12):6589-6603. doi: 10.21037/jtd-23-653. Epub 2023 Dec 15. J Thorac Dis. 2023. PMID: 38249879 Free PMC article.
-
ICU admission and mortality classifiers for COVID-19 patients based on subgroups of dynamically associated profiles across multiple timepoints.Comput Biol Med. 2022 Feb;141:105176. doi: 10.1016/j.compbiomed.2021.105176. Epub 2021 Dec 27. Comput Biol Med. 2022. PMID: 35007991 Free PMC article.
-
DeepASD: a deep adversarial-regularized graph learning method for ASD diagnosis with multimodal data.Transl Psychiatry. 2024 Sep 14;14(1):375. doi: 10.1038/s41398-024-02972-2. Transl Psychiatry. 2024. PMID: 39277595 Free PMC article.
-
Machine learning and predictive models: 2 years of Sars-CoV-2 pandemic in a single-center retrospective analysis.J Anesth Analg Crit Care. 2022 Oct 14;2(1):42. doi: 10.1186/s44158-022-00071-6. J Anesth Analg Crit Care. 2022. PMID: 37386654 Free PMC article.
-
COVID-19 Diagnosis from Chest X-ray Images Using a Robust Multi-Resolution Analysis Siamese Neural Network with Super-Resolution Convolutional Neural Network.Diagnostics (Basel). 2022 Mar 18;12(3):741. doi: 10.3390/diagnostics12030741. Diagnostics (Basel). 2022. PMID: 35328294 Free PMC article.
References
-
- Weekly epidemiological update - 12 January 2021. World Health Organization. 2021. Jan 12, [2021-03-18]. https://www.who.int/publications/m/item/weekly-epidemiological-update---....
-
- Wölfel Roman, Corman Victor M, Guggemos Wolfgang, Seilmaier Michael, Zange Sabine, Müller Marcel A, Niemeyer Daniela, Jones Terry C, Vollmar Patrick, Rothe Camilla, Hoelscher Michael, Bleicker Tobias, Brünink Sebastian, Schneider Julia, Ehmann Rosina, Zwirglmaier Katrin, Drosten Christian, Wendtner Clemens. Virological assessment of hospitalized patients with COVID-2019. Nature. 2020 May;581(7809):465–469. doi: 10.1038/s41586-020-2196-x. doi: 10.1038/s41586-020-2196-x. - DOI - DOI - PubMed
-
- Luo Yi, Trevathan Edwin, Qian Zhengmin, Li Yirong, Li Jin, Xiao Wei, Tu Ning, Zeng Zhikun, Mo Pingzheng, Xiong Yong, Ye Guangming. Asymptomatic SARS-CoV-2 Infection in Household Contacts of a Healthcare Provider, Wuhan, China. Emerg Infect Dis. 2020 Aug;26(8):1930–1933. doi: 10.3201/eid2608.201016. doi: 10.3201/eid2608.201016. - DOI - DOI - PMC - PubMed
-
- Ye F, Xu S, Rong Z, Xu R, Liu X, Deng P, Liu H, Xu X. Delivery of infection from asymptomatic carriers of COVID-19 in a familial cluster. Int J Infect Dis. 2020 May;94:133–138. doi: 10.1016/j.ijid.2020.03.042. https://linkinghub.elsevier.com/retrieve/pii/S1201-9712(20)30174-0 - DOI - PMC - PubMed
Publication types
MeSH terms
LinkOut - more resources
Full Text Sources
Medical
Miscellaneous