Cargando…

Hybrid Text Feature Modeling for Disease Group Prediction Using Unstructured Physician Notes

Existing Clinical Decision Support Systems (CDSSs) largely depend on the availability of structured patient data and Electronic Health Records (EHRs) to aid caregivers. However, in case of hospitals in developing countries, structured patient data formats are not widely adopted, where medical profes...

Descripción completa

Detalles Bibliográficos
Autores principales: Krishnan, Gokul S., Kamath, S. Sowmya
Formato: Online Artículo Texto
Lenguaje:English
Publicado: 2020
Materias:
Acceso en línea:https://www.ncbi.nlm.nih.gov/pmc/articles/PMC7303721/
http://dx.doi.org/10.1007/978-3-030-50423-6_24
_version_ 1783548120857051136
author Krishnan, Gokul S.
Kamath, S. Sowmya
author_facet Krishnan, Gokul S.
Kamath, S. Sowmya
author_sort Krishnan, Gokul S.
collection PubMed
description Existing Clinical Decision Support Systems (CDSSs) largely depend on the availability of structured patient data and Electronic Health Records (EHRs) to aid caregivers. However, in case of hospitals in developing countries, structured patient data formats are not widely adopted, where medical professionals still rely on clinical notes in the form of unstructured text. Such unstructured clinical notes recorded by medical personnel can also be a potential source of rich patient-specific information which can be leveraged to build CDSSs, even for hospitals in developing countries. If such unstructured clinical text can be used, the manual and time-consuming process of EHR generation will no longer be required, with huge person-hours and cost savings. In this article, we propose a generic ICD9 disease group prediction CDSS built on unstructured physician notes modeled using hybrid word embeddings. These word embeddings are used to train a deep neural network for effectively predicting ICD9 disease groups. Experimental evaluation showed that the proposed approach outperformed the state-of-the-art disease group prediction model built on structured EHRs by 15% in terms of AUROC and 40% in terms of AUPRC, thus proving our hypothesis and eliminating dependency on availability of structured patient data.
format Online
Article
Text
id pubmed-7303721
institution National Center for Biotechnology Information
language English
publishDate 2020
record_format MEDLINE/PubMed
spelling pubmed-73037212020-06-19 Hybrid Text Feature Modeling for Disease Group Prediction Using Unstructured Physician Notes Krishnan, Gokul S. Kamath, S. Sowmya Computational Science – ICCS 2020 Article Existing Clinical Decision Support Systems (CDSSs) largely depend on the availability of structured patient data and Electronic Health Records (EHRs) to aid caregivers. However, in case of hospitals in developing countries, structured patient data formats are not widely adopted, where medical professionals still rely on clinical notes in the form of unstructured text. Such unstructured clinical notes recorded by medical personnel can also be a potential source of rich patient-specific information which can be leveraged to build CDSSs, even for hospitals in developing countries. If such unstructured clinical text can be used, the manual and time-consuming process of EHR generation will no longer be required, with huge person-hours and cost savings. In this article, we propose a generic ICD9 disease group prediction CDSS built on unstructured physician notes modeled using hybrid word embeddings. These word embeddings are used to train a deep neural network for effectively predicting ICD9 disease groups. Experimental evaluation showed that the proposed approach outperformed the state-of-the-art disease group prediction model built on structured EHRs by 15% in terms of AUROC and 40% in terms of AUPRC, thus proving our hypothesis and eliminating dependency on availability of structured patient data. 2020-05-23 /pmc/articles/PMC7303721/ http://dx.doi.org/10.1007/978-3-030-50423-6_24 Text en © Springer Nature Switzerland AG 2020 This article is made available via the PMC Open Access Subset for unrestricted research re-use and secondary analysis in any form or by any means with acknowledgement of the original source. These permissions are granted for the duration of the World Health Organization (WHO) declaration of COVID-19 as a global pandemic.
spellingShingle Article
Krishnan, Gokul S.
Kamath, S. Sowmya
Hybrid Text Feature Modeling for Disease Group Prediction Using Unstructured Physician Notes
title Hybrid Text Feature Modeling for Disease Group Prediction Using Unstructured Physician Notes
title_full Hybrid Text Feature Modeling for Disease Group Prediction Using Unstructured Physician Notes
title_fullStr Hybrid Text Feature Modeling for Disease Group Prediction Using Unstructured Physician Notes
title_full_unstemmed Hybrid Text Feature Modeling for Disease Group Prediction Using Unstructured Physician Notes
title_short Hybrid Text Feature Modeling for Disease Group Prediction Using Unstructured Physician Notes
title_sort hybrid text feature modeling for disease group prediction using unstructured physician notes
topic Article
url https://www.ncbi.nlm.nih.gov/pmc/articles/PMC7303721/
http://dx.doi.org/10.1007/978-3-030-50423-6_24
work_keys_str_mv AT krishnangokuls hybridtextfeaturemodelingfordiseasegrouppredictionusingunstructuredphysiciannotes
AT kamathssowmya hybridtextfeaturemodelingfordiseasegrouppredictionusingunstructuredphysiciannotes