Natural language processing to identify lupus nephritis phenotype in electronic health records

Yu Deng,Jennifer A. Pacheco,Anika Ghosh,Anh Chung,Chengsheng Mao,Joshua C. Smith,Juan Zhao,Wei-Qi Wei,April Barnado,Chad Dorn,Chunhua Weng,Cong Liu,Adam Cordon,Jingzhi Yu,Yacob Tedla,Abel Kho,Rosalind Ramsey-Goldman,Theresa Walunas,Yuan Luo
DOI: https://doi.org/10.1186/s12911-024-02420-7
IF: 3.298
2024-03-04
BMC Medical Informatics and Decision Making
Abstract:Systemic lupus erythematosus (SLE) is a rare autoimmune disorder characterized by an unpredictable course of flares and remission with diverse manifestations. Lupus nephritis, one of the major disease manifestations of SLE for organ damage and mortality, is a key component of lupus classification criteria. Accurately identifying lupus nephritis in electronic health records (EHRs) would therefore benefit large cohort observational studies and clinical trials where characterization of the patient population is critical for recruitment, study design, and analysis. Lupus nephritis can be recognized through procedure codes and structured data, such as laboratory tests. However, other critical information documenting lupus nephritis, such as histologic reports from kidney biopsies and prior medical history narratives, require sophisticated text processing to mine information from pathology reports and clinical notes. In this study, we developed algorithms to identify lupus nephritis with and without natural language processing (NLP) using EHR data from the Northwestern Medicine Enterprise Data Warehouse (NMEDW).
medical informatics
What problem does this paper attempt to address?