Back to Search Start Over

Clinical named entity recognition and relation extraction using natural language processing of medical free text: A systematic review.

Authors :
Fraile Navarro D
Ijaz K
Rezazadegan D
Rahimi-Ardabili H
Dras M
Coiera E
Berkovsky S
Source :
International journal of medical informatics [Int J Med Inform] 2023 Sep; Vol. 177, pp. 105122. Date of Electronic Publication: 2023 Jun 05.
Publication Year :
2023

Abstract

Background: Natural Language Processing (NLP) applications have developed over the past years in various fields including its application to clinical free text for named entity recognition and relation extraction. However, there has been rapid developments the last few years that there's currently no overview of it. Moreover, it is unclear how these models and tools have been translated into clinical practice. We aim to synthesize and review these developments.<br />Methods: We reviewed literature from 2010 to date, searching PubMed, Scopus, the Association of Computational Linguistics (ACL), and Association of Computer Machinery (ACM) libraries for studies of NLP systems performing general-purpose (i.e., not disease- or treatment-specific) information extraction and relation extraction tasks in unstructured clinical text (e.g., discharge summaries).<br />Results: We included in the review 94 studies with 30 studies published in the last three years. Machine learning methods were used in 68 studies, rule-based in 5 studies, and both in 22 studies. 63 studies focused on Named Entity Recognition, 13 on Relation Extraction and 18 performed both. The most frequently extracted entities were "problem", "test" and "treatment". 72 studies used public datasets and 22 studies used proprietary datasets alone. Only 14 studies defined clearly a clinical or information task to be addressed by the system and just three studies reported its use outside the experimental setting. Only 7 studies shared a pre-trained model and only 8 an available software tool.<br />Discussion: Machine learning-based methods have dominated the NLP field on information extraction tasks. More recently, Transformer-based language models are taking the lead and showing the strongest performance. However, these developments are mostly based on a few datasets and generic annotations, with very few real-world use cases. This may raise questions about the generalizability of findings, translation into practice and highlights the need for robust clinical evaluation.<br />Competing Interests: Declaration of Competing Interest The authors declare that they have no known competing financial interests or personal relationships that could have appeared to influence the work reported in this paper.<br /> (Copyright © 2023 Elsevier B.V. All rights reserved.)

Details

Language :
English
ISSN :
1872-8243
Volume :
177
Database :
MEDLINE
Journal :
International journal of medical informatics
Publication Type :
Academic Journal
Accession number :
37295138
Full Text :
https://doi.org/10.1016/j.ijmedinf.2023.105122