Information extraction from German radiological reports for general clinical text and language understanding

Autor:	Michael Jantscher, Felix Gunzer, Roman Kern, Eva Hassler, Sebastian Tschauner, Gernot Reishofer
Jazyk:	angličtina
Rok vydání:	2023
Předmět:	Medicine Science
Zdroj:	Scientific Reports, Vol 13, Iss 1, Pp 1-12 (2023)
Druh dokumentu:	article
ISSN:	2045-2322
DOI:	10.1038/s41598-023-29323-3
Popis:	Abstract Recent advances in deep learning and natural language processing (NLP) have opened many new opportunities for automatic text understanding and text processing in the medical field. This is of great benefit as many clinical downstream tasks rely on information from unstructured clinical documents. However, for low-resource languages like German, the use of modern text processing applications that require a large amount of training data proves to be difficult, as only few data sets are available mainly due to legal restrictions. In this study, we present an information extraction framework that was initially pre-trained on real-world computed tomographic (CT) reports of head examinations, followed by domain adaptive fine-tuning on reports from different imaging examinations. We show that in the pre-training phase, the semantic and contextual meaning of one clinical reporting domain can be captured and effectively transferred to foreign clinical imaging examinations. Moreover, we introduce an active learning approach with an intrinsic strategic sampling method to generate highly informative training data with low human annotation cost. We see that the model performance can be significantly improved by an appropriate selection of the data to be annotated, without the need to train the model on a specific downstream task. With a general annotation scheme that can be used not only in the radiology field but also in a broader clinical setting, we contribute to a more consistent labeling and annotation process that also facilitates the verification and evaluation of language models in the German clinical setting.
Databáze:	Directory of Open Access Journals
Externí odkaz:	https://doaj.org/article/8c6c21555e5341928d2e3ac1adc860f0 Zobrazit plný text záznamu View record in DOAJ Plný text ve formátu PDF Plný text ve formátu HTML
Nepřihlášeným uživatelům se plný text nezobrazuje	K zobrazení výsledku je třeba se přihlásit.