Author
Abstract
The emergence of COVID-19 in late 2019 in Wuhan, China, has led to a global health crisis that has claimed many lives worldwide. A thorough understanding of the available COVID-19 datasets can enable healthcare professionals to identify cases at an early stage. This study presents an innovative pipeline-based framework for predicting survival and mortality in patients with COVID-19 by leveraging the Mexican COVID-19 patient dataset (COVID-19-MPD dataset). Preprocessing plays a pivotal role in ensuring that the framework delivers high-quality outcomes. We deploy various machine learning models with optimized hyperparameters within the framework. Through consistent experimental conditions and dataset utilization, we conducted multiple experiments employing diverse preprocessing techniques and models to maximize the area under the receiver operating characteristic curve (AUC) for COVID-19 prediction. Given the considerable dimensions of the dataset, feature selection is crucial for identifying factors influencing COVID-19 mortality or survival. We employ feature dimension reduction methods, such as principal component analysis and independent component analysis, in addition to feature selection techniques such as maximum relevance minimum redundancy and permutation feature importance. Impactful features related to patient outcomes can significantly aid experts in disease management by enhancing treatment efficacy and control measures. Following various experiments with standardized data and AUC assessment using the k-nearest neighbor algorithm with four components, the proposed framework achieves optimal results, attaining an AUC of 100%. Given its effectiveness in COVID-19 prediction, this framework has the potential for integration into medical decision support systems. Graphical abstract
Suggested Citation
Rahman Farnoosh & Karlo Abnoosian, 2024.
"A robust innovative pipeline-based machine learning framework for predicting COVID-19 in Mexican patients,"
International Journal of System Assurance Engineering and Management, Springer;The Society for Reliability, Engineering Quality and Operations Management (SREQOM),India, and Division of Operation and Maintenance, Lulea University of Technology, Sweden, vol. 15(7), pages 3466-3484, July.
Handle:
RePEc:spr:ijsaem:v:15:y:2024:i:7:d:10.1007_s13198-024-02354-3
DOI: 10.1007/s13198-024-02354-3
Download full text from publisher
As the access to this document is restricted, you may want to search for a different version of it.
Corrections
All material on this site has been provided by the respective publishers and authors. You can help correct errors and omissions. When requesting a correction, please mention this item's handle: RePEc:spr:ijsaem:v:15:y:2024:i:7:d:10.1007_s13198-024-02354-3. See general information about how to correct material in RePEc.
If you have authored this item and are not yet registered with RePEc, we encourage you to do it here. This allows to link your profile to this item. It also allows you to accept potential citations to this item that we are uncertain about.
We have no bibliographic references for this item. You can help adding them by using this form .
If you know of missing items citing this one, you can help us creating those links by adding the relevant references in the same way as above, for each refering item. If you are a registered author of this item, you may also want to check the "citations" tab in your RePEc Author Service profile, as there may be some citations waiting for confirmation.
For technical questions regarding this item, or to correct its authors, title, abstract, bibliographic or download information, contact: Sonal Shukla or Springer Nature Abstracting and Indexing (email available below). General contact details of provider: http://www.springer.com .
Please note that corrections may take a couple of weeks to filter through
the various RePEc services.