Natural language processing of MIMIC-III clinical notes for identifying diagnosis and procedures with neural networks

Nuthakki, Siddhartha; Neela, Sunil; Gichoya, Judy W.; Purkayastha, Saptarshi

Natural language processing of MIMIC-III clinical notes for identifying diagnosis and procedures with neural networks

dc.contributor.author	Nuthakki, Siddhartha
dc.contributor.author	Neela, Sunil
dc.contributor.author	Gichoya, Judy W.
dc.contributor.author	Purkayastha, Saptarshi
dc.contributor.department	BioHealth Informatics, School of Informatics and Computing	en_US
dc.date.accessioned	2022-10-06T19:14:16Z
dc.date.available	2022-10-06T19:14:16Z
dc.date.issued	2019
dc.description.abstract	Coding diagnosis and procedures in medical records is a crucial process in the healthcare industry, which includes the creation of accurate billings, receiving reimbursements from payers, and creating standardized patient care records. In the United States, Billing and Insurance related activities cost around $471 billion in 2012 which constitutes about 25% of all the U.S hospital spending. In this paper, we report the performance of a natural language processing model that can map clinical notes to medical codes, and predict final diagnosis from unstructured entries of history of present illness, symptoms at the time of admission, etc. Previous studies have demonstrated that deep learning models perform better at such mapping when compared to conventional machine learning models. Therefore, we employed state-of-the-art deep learning method, ULMFiT on the largest emergency department clinical notes dataset MIMIC III which has 1.2M clinical notes to select for the top-10 and top-50 diagnosis and procedure codes. Our models were able to predict the top-10 diagnoses and procedures with 80.3% and 80.5% accuracy, whereas the top-50 ICD-9 codes of diagnosis and procedures are predicted with 70.7% and 63.9% accuracy. Prediction of diagnosis and procedures from unstructured clinical notes benefit human coders to save time, eliminate errors and minimize costs. With promising scores from our present model, the next step would be to deploy this on a small-scale real-world scenario and compare it with human coders as the gold standard. We believe that further research of this approach can create highly accurate predictions that can ease the workflow in a clinical setting.	en_US
dc.eprint.version	Author's manuscript	en_US
dc.identifier.citation	Nuthakki, S., Neela, S., Gichoya, J. W., & Purkayastha, S. (2019). Natural language processing of MIMIC-III clinical notes for identifying diagnosis and procedures with neural networks. arXiv preprint arXiv:1912.12397. https://doi.org/10.48550/arXiv.1912.12397	en_US
dc.identifier.uri	https://hdl.handle.net/1805/30240
dc.language.iso	en	en_US
dc.publisher	arXiv	en_US
dc.relation.isversionof	10.48550/arXiv.1912.12397	en_US
dc.relation.journal	arXiv	en_US
dc.rights	IUPUI Open Access Policy	en_US
dc.source	Author	en_US
dc.subject	natural language processing	en_US
dc.subject	clinical notes	en_US
dc.subject	neural networks	en_US
dc.title	Natural language processing of MIMIC-III clinical notes for identifying diagnosis and procedures with neural networks	en_US
dc.type	Article	en_US

Files

Original bundle

Now showing 1 - 1 of 1

Name:: Nuthakki2019Natural-preprint.pdf
Size:: 838.53 KB
Format:: Adobe Portable Document Format
Description:

Download

License bundle

Now showing 1 - 1 of 1

Name:: license.txt
Size:: 1.99 KB
Format:: Item-specific license agreed upon to submission
Description:

Download

Collections

Open Access Policy Articles
Department of Biomedical Engineering and Informatics Works
Saptarshi Purkayastha