Your browser doesn't support javascript.
loading
Automated System to Capture Patient Symptoms From Multitype Japanese Clinical Texts: Retrospective Study.
Nishiyama, Tomohiro; Yamaguchi, Ayane; Han, Peitao; Pereira, Lis Weiji Kanashiro; Otsuki, Yuka; Andrade, Gabriel Herman Bernardim; Kudo, Noriko; Yada, Shuntaro; Wakamiya, Shoko; Aramaki, Eiji; Takada, Masahiro; Toi, Masakazu.
Affiliation
  • Nishiyama T; Department of Information Science, Nara Institute of Science and Technology, Ikoma, Japan.
  • Yamaguchi A; Graduate School of Medicine, Kyoto University, Kyoto, Japan.
  • Han P; Department of Information Science, Nara Institute of Science and Technology, Ikoma, Japan.
  • Pereira LWK; Center for Information and Neural Networks, Advanced ICT Research Institute, Osaka, Japan.
  • Otsuki Y; Department of Information Science, Nara Institute of Science and Technology, Ikoma, Japan.
  • Andrade GHB; Department of Information Science, Nara Institute of Science and Technology, Ikoma, Japan.
  • Kudo N; Department of Information Science, Nara Institute of Science and Technology, Ikoma, Japan.
  • Yada S; Department of Information Science, Nara Institute of Science and Technology, Ikoma, Japan.
  • Wakamiya S; Department of Information Science, Nara Institute of Science and Technology, Ikoma, Japan.
  • Aramaki E; Department of Information Science, Nara Institute of Science and Technology, Ikoma, Japan.
  • Takada M; Graduate School of Medicine, Kyoto University, Kyoto, Japan.
  • Toi M; Department of Breast Surgery, Kansai Medical University, Hirakata, Japan.
JMIR Med Inform ; 12: e58977, 2024 Sep 24.
Article in En | MEDLINE | ID: mdl-39316418
ABSTRACT

BACKGROUND:

Natural language processing (NLP) techniques can be used to analyze large amounts of electronic health record texts, which encompasses various types of patient information such as quality of life, effectiveness of treatments, and adverse drug event (ADE) signals. As different aspects of a patient's status are stored in different types of documents, we propose an NLP system capable of processing 6 types of documents physician progress notes, discharge summaries, radiology reports, radioisotope reports, nursing records, and pharmacist progress notes.

OBJECTIVE:

This study aimed to investigate the system's performance in detecting ADEs by evaluating the results from multitype texts. The main objective is to detect adverse events accurately using an NLP system.

METHODS:

We used data written in Japanese from 2289 patients with breast cancer, including medication data, physician progress notes, discharge summaries, radiology reports, radioisotope reports, nursing records, and pharmacist progress notes. Our system performs 3 processes named entity recognition, normalization of symptoms, and aggregation of multiple types of documents from multiple patients. Among all patients with breast cancer, 103 and 112 with peripheral neuropathy (PN) received paclitaxel or docetaxel, respectively. We evaluate the utility of using multiple types of documents by correlation coefficient and regression analysis to compare their performance with each single type of document. All evaluations of detection rates with our system are performed 30 days after drug administration.

RESULTS:

Our system underestimates by 13.3 percentage points (74.0%-60.7%), as the incidence of paclitaxel-induced PN was 60.7%, compared with 74.0% in the previous research based on manual extraction. The Pearson correlation coefficient between the manual extraction and system results was 0.87 Although the pharmacist progress notes had the highest detection rate among each type of document, the rate did not match the performance using all documents. The estimated median duration of PN with paclitaxel was 92 days, whereas the previously reported median duration of PN with paclitaxel was 727 days. The number of events detected in each document was highest in the physician's progress notes, followed by the pharmacist's and nursing records.

CONCLUSIONS:

Considering the inherent cost that requires constant monitoring of the patient's condition, such as the treatment of PN, our system has a significant advantage in that it can immediately estimate the treatment duration without fine-tuning a new NLP model. Leveraging multitype documents is better than using single-type documents to improve detection performance. Although the onset time estimation was relatively accurate, the duration might have been influenced by the length of the data follow-up period. The results suggest that our method using various types of data can detect more ADEs from clinical documents.
Subject(s)
Key words

Full text: 1 Collection: 01-internacional Database: MEDLINE Main subject: Natural Language Processing / Electronic Health Records Limits: Female / Humans Country/Region as subject: Asia Language: En Journal: JMIR Med Inform Year: 2024 Document type: Article Affiliation country: Japan Country of publication: Canada

Full text: 1 Collection: 01-internacional Database: MEDLINE Main subject: Natural Language Processing / Electronic Health Records Limits: Female / Humans Country/Region as subject: Asia Language: En Journal: JMIR Med Inform Year: 2024 Document type: Article Affiliation country: Japan Country of publication: Canada