Cargando…

A machine learning model on Real World Data for predicting progression to Acute Respiratory Distress Syndrome (ARDS) among COVID-19 patients

INTRODUCTION: Identifying COVID-19 patients that are most likely to progress to a severe infection is crucial for optimizing care management and increasing the likelihood of survival. This study presents a machine learning model that predicts severe cases of COVID-19, defined as the presence of Acut...

Descripción completa

Detalles Bibliográficos
Autores principales: Lazzarini, Nicola, Filippoupolitis, Avgoustinos, Manzione, Pedro, Eleftherohorinou, Hariklia
Formato: Online Artículo Texto
Lenguaje:English
Publicado: Public Library of Science 2022
Materias:
Acceso en línea:https://www.ncbi.nlm.nih.gov/pmc/articles/PMC9333235/
https://www.ncbi.nlm.nih.gov/pubmed/35901089
http://dx.doi.org/10.1371/journal.pone.0271227
_version_ 1784758828052512768
author Lazzarini, Nicola
Filippoupolitis, Avgoustinos
Manzione, Pedro
Eleftherohorinou, Hariklia
author_facet Lazzarini, Nicola
Filippoupolitis, Avgoustinos
Manzione, Pedro
Eleftherohorinou, Hariklia
author_sort Lazzarini, Nicola
collection PubMed
description INTRODUCTION: Identifying COVID-19 patients that are most likely to progress to a severe infection is crucial for optimizing care management and increasing the likelihood of survival. This study presents a machine learning model that predicts severe cases of COVID-19, defined as the presence of Acute Respiratory Distress Syndrome (ARDS) and highlights the different risk factors that play a significant role in disease progression. METHODS: A cohort composed of 289,351 patients diagnosed with COVID-19 in April 2020 was created using US administrative claims data from Oct 2015 to Jul 2020. For each patient, information about 817 diagnoses, were collected from the medical history ahead of COVID-19 infection. The primary outcome of the study was the presence of ARDS in the 4 months following COVID-19 infection. The study cohort was randomly split into training set used for model development, test set for model evaluation and validation set for real-world performance estimation. RESULTS: We analyzed three machine learning classifiers to predict the presence of ARDS. Among the algorithms considered, a Gradient Boosting Decision Tree had the highest performance with an AUC of 0.695 (95% CI, 0.679–0.709) and an AUPRC of 0.0730 (95% CI, 0.0676 – 0.0823), showing a 40% performance increase in performance against a baseline classifier. A panel of five clinicians was also used to compare the predictive ability of the model to that of clinical experts. The comparison indicated that our model is on par or outperforms predictions made by the clinicians, both in terms of precision and recall. CONCLUSION: This study presents a machine learning model that uses patient claims history to predict ARDS. The risk factors used by the model to perform its predictions have been extensively linked to the severity of the COVID-19 in the specialized literature. The most contributing diagnosis can be easily retrieved in the patient clinical history and can be used for an early screening of infected patients. Overall, the proposed model could be a promising tool to deploy in a healthcare setting to facilitate and optimize the care of COVID-19 patients.
format Online
Article
Text
id pubmed-9333235
institution National Center for Biotechnology Information
language English
publishDate 2022
publisher Public Library of Science
record_format MEDLINE/PubMed
spelling pubmed-93332352022-07-29 A machine learning model on Real World Data for predicting progression to Acute Respiratory Distress Syndrome (ARDS) among COVID-19 patients Lazzarini, Nicola Filippoupolitis, Avgoustinos Manzione, Pedro Eleftherohorinou, Hariklia PLoS One Research Article INTRODUCTION: Identifying COVID-19 patients that are most likely to progress to a severe infection is crucial for optimizing care management and increasing the likelihood of survival. This study presents a machine learning model that predicts severe cases of COVID-19, defined as the presence of Acute Respiratory Distress Syndrome (ARDS) and highlights the different risk factors that play a significant role in disease progression. METHODS: A cohort composed of 289,351 patients diagnosed with COVID-19 in April 2020 was created using US administrative claims data from Oct 2015 to Jul 2020. For each patient, information about 817 diagnoses, were collected from the medical history ahead of COVID-19 infection. The primary outcome of the study was the presence of ARDS in the 4 months following COVID-19 infection. The study cohort was randomly split into training set used for model development, test set for model evaluation and validation set for real-world performance estimation. RESULTS: We analyzed three machine learning classifiers to predict the presence of ARDS. Among the algorithms considered, a Gradient Boosting Decision Tree had the highest performance with an AUC of 0.695 (95% CI, 0.679–0.709) and an AUPRC of 0.0730 (95% CI, 0.0676 – 0.0823), showing a 40% performance increase in performance against a baseline classifier. A panel of five clinicians was also used to compare the predictive ability of the model to that of clinical experts. The comparison indicated that our model is on par or outperforms predictions made by the clinicians, both in terms of precision and recall. CONCLUSION: This study presents a machine learning model that uses patient claims history to predict ARDS. The risk factors used by the model to perform its predictions have been extensively linked to the severity of the COVID-19 in the specialized literature. The most contributing diagnosis can be easily retrieved in the patient clinical history and can be used for an early screening of infected patients. Overall, the proposed model could be a promising tool to deploy in a healthcare setting to facilitate and optimize the care of COVID-19 patients. Public Library of Science 2022-07-28 /pmc/articles/PMC9333235/ /pubmed/35901089 http://dx.doi.org/10.1371/journal.pone.0271227 Text en © 2022 Lazzarini et al https://creativecommons.org/licenses/by/4.0/This is an open access article distributed under the terms of the Creative Commons Attribution License (https://creativecommons.org/licenses/by/4.0/) , which permits unrestricted use, distribution, and reproduction in any medium, provided the original author and source are credited.
spellingShingle Research Article
Lazzarini, Nicola
Filippoupolitis, Avgoustinos
Manzione, Pedro
Eleftherohorinou, Hariklia
A machine learning model on Real World Data for predicting progression to Acute Respiratory Distress Syndrome (ARDS) among COVID-19 patients
title A machine learning model on Real World Data for predicting progression to Acute Respiratory Distress Syndrome (ARDS) among COVID-19 patients
title_full A machine learning model on Real World Data for predicting progression to Acute Respiratory Distress Syndrome (ARDS) among COVID-19 patients
title_fullStr A machine learning model on Real World Data for predicting progression to Acute Respiratory Distress Syndrome (ARDS) among COVID-19 patients
title_full_unstemmed A machine learning model on Real World Data for predicting progression to Acute Respiratory Distress Syndrome (ARDS) among COVID-19 patients
title_short A machine learning model on Real World Data for predicting progression to Acute Respiratory Distress Syndrome (ARDS) among COVID-19 patients
title_sort machine learning model on real world data for predicting progression to acute respiratory distress syndrome (ards) among covid-19 patients
topic Research Article
url https://www.ncbi.nlm.nih.gov/pmc/articles/PMC9333235/
https://www.ncbi.nlm.nih.gov/pubmed/35901089
http://dx.doi.org/10.1371/journal.pone.0271227
work_keys_str_mv AT lazzarininicola amachinelearningmodelonrealworlddataforpredictingprogressiontoacuterespiratorydistresssyndromeardsamongcovid19patients
AT filippoupolitisavgoustinos amachinelearningmodelonrealworlddataforpredictingprogressiontoacuterespiratorydistresssyndromeardsamongcovid19patients
AT manzionepedro amachinelearningmodelonrealworlddataforpredictingprogressiontoacuterespiratorydistresssyndromeardsamongcovid19patients
AT eleftherohorinouhariklia amachinelearningmodelonrealworlddataforpredictingprogressiontoacuterespiratorydistresssyndromeardsamongcovid19patients
AT lazzarininicola machinelearningmodelonrealworlddataforpredictingprogressiontoacuterespiratorydistresssyndromeardsamongcovid19patients
AT filippoupolitisavgoustinos machinelearningmodelonrealworlddataforpredictingprogressiontoacuterespiratorydistresssyndromeardsamongcovid19patients
AT manzionepedro machinelearningmodelonrealworlddataforpredictingprogressiontoacuterespiratorydistresssyndromeardsamongcovid19patients
AT eleftherohorinouhariklia machinelearningmodelonrealworlddataforpredictingprogressiontoacuterespiratorydistresssyndromeardsamongcovid19patients