Adapting State-of-the-Art Deep Language Models to Clinical Information Extraction Systems: Potentials, Challenges, and Solutions

Tom Gedeon; Liyuan Zhou; Hanna Suominen

Adapting State-of-the-Art Deep Language Models to Clinical Information Extraction Systems: Potentials, Challenges, and Solutions

Tom Gedeon; Liyuan Zhou; Hanna Suominen

dc.contributor.author	Tom Gedeon
dc.contributor.author	Liyuan Zhou
dc.contributor.author	Hanna Suominen
dc.date.accessioned	2022-10-28T13:32:11Z
dc.date.available	2022-10-28T13:32:11Z
dc.identifier.uri	https://www.utupub.fi/handle/10024/165850
dc.description.abstract	<p>Background: Deep learning (DL) has been widely used to solve problems with success in speech recognition, visual object recognition, and object detection for drug discovery and genomics. Natural language processing has achieved noticeable progress in artificial intelligence. This gives an opportunity to improve on the accuracy and human-computer interaction of clinical informatics. However, due to difference of vocabularies and context between a clinical environment and generic English, transplanting language models directly from up-to-date methods to real-world health care settings is not always satisfactory. Moreover, the legal restriction on using privacy-sensitive patient records hinders the progress in applying machine learning (ML) to clinical language processing.</p><p>Objective: The aim of this study was to investigate 2 ways to adapt state-of-the-art language models to extracting patient information from free-form clinical narratives to populate a handover form at a nursing shift change automatically for proofing and revising by hand: first, by using domain-specific word representations and second, by using transfer learning models to adapt knowledge from general to clinical English. We have described the practical problem, composed it as an ML task known as information extraction, proposed methods for solving the task, and evaluated their performance.</p><p>Methods: First, word representations trained from different domains served as the input of a DL system for information extraction. Second, the transfer learning model was applied as a way to adapt the knowledge learned from general text sources to the task domain. The goal was to gain improvements in the extraction performance, especially for the classes that were topically related but did not have a sufficient amount of model solutions available for ML directly from the target domain. A total of 3 independent datasets were generated for this task, and they were used as the training (101 patient reports), validation (100 patient reports), and test (100 patient reports) sets in our experiments.</p><p>Results: Our system is now the state-of-the-art in this task. Domain-specific word representations improved the macroaveraged F1 by 3.4%. Transferring the knowledge from general English corpora to the task-specific domain contributed a further 7.1% improvement. The best performance in populating the handover form with 37 headings was the macroaveraged F1 of 41.6% and F1 of 81.1% for filtering out irrelevant information. Performance differences between this system and its baseline were statistically significant (P<.001; Wilcoxon test).</p><p>Conclusions: To our knowledge, our study is the first attempt to transfer models from general deep models to specific tasks in health care and gain a significant improvement. As transfer learning shows its advantage over other methods, especially on classes with a limited amount of training data, less experts' time is needed to annotate data for ML, which may enable good results even in resource-poor domains.</p>
dc.language.iso	en
dc.publisher	JMIR PUBLICATIONS, INC
dc.title	Adapting State-of-the-Art Deep Language Models to Clinical Information Extraction Systems: Potentials, Challenges, and Solutions
dc.identifier.url	https://medinform.jmir.org/2019/2/e11499/
dc.identifier.urn	URN:NBN:fi-fe2021042827531
dc.relation.volume	7
dc.contributor.organization	fi=PÄÄT Tulevaisuuden teknologioiden laitos, yhteiset\|en=PÄÄT Department of Future Technologies\|
dc.contributor.organization-code	2606800
dc.converis.publication-id	41729663
dc.converis.url	https://research.utu.fi/converis/portal/Publication/41729663
dc.format.pagerange	81
dc.format.pagerange	67
dc.identifier.eissn	2291-9694
dc.identifier.jour-issn	2291-9694
dc.okm.affiliatedauthor	Suominen, Hanna
dc.okm.discipline	113 Tietojenkäsittely ja informaatiotieteet	fi_FI
dc.okm.discipline	113 Computer and information sciences	en_GB
dc.okm.internationalcopublication	international co-publication
dc.okm.internationality	International publication
dc.okm.type	Journal article
dc.publisher.country	Canada	en_GB
dc.publisher.country	Kanada	fi_FI
dc.publisher.country-code	CA
dc.relation.articlenumber	ARTN e11499
dc.relation.doi	10.2196/11499
dc.relation.ispartofjournal	JMIR Medical Informatics
dc.relation.issue	2
dc.year.issued	2019

Aineistoon kuuluvat tiedostot

Nimi:: SuominenEtAl AdaptingState.pdf
Koko:: 316.9Kb
Tiedostomuoto:: PDF
Kuvaus:: Publisher's pdf

Katso/Avaa

Aineisto kuuluu seuraaviin kokoelmiin

Rinnakkaistallenteet [19207]

Näytä suppeat kuvailutiedot