• Home
  • Search
  • Automatic detection of actionable radiology reports using bidirectional encoder representations from transformers
  • Cite Icon38
  • https://doi.org/10.1186/s12911-021-01623-6Copy DOI Icon

Automatic detection of actionable radiology reports using bidirectional encoder representations from transformers

Show More
  • Abstract
  • Highlights & Summary
  • PDF
  • Literature Map
  • References
  • Citations
  • Similar Papers
Abstract

BackgroundIt is essential for radiologists to communicate actionable findings to the referring clinicians reliably. Natural language processing (NLP) has been shown to help identify free-text radiology reports including actionable findings. However, the application of recent deep learning techniques to radiology reports, which can improve the detection performance, has not been thoroughly examined. Moreover, free-text that clinicians input in the ordering form (order information) has seldom been used to identify actionable reports. This study aims to evaluate the benefits of two new approaches: (1) bidirectional encoder representations from transformers (BERT), a recent deep learning architecture in NLP, and (2) using order information in addition to radiology reports.MethodsWe performed a binary classification to distinguish actionable reports (i.e., radiology reports tagged as actionable in actual radiological practice) from non-actionable ones (those without an actionable tag). 90,923 Japanese radiology reports in our hospital were used, of which 788 (0.87%) were actionable. We evaluated four methods, statistical machine learning with logistic regression (LR) and with gradient boosting decision tree (GBDT), and deep learning with a bidirectional long short-term memory (LSTM) model and a publicly available Japanese BERT model. Each method was used with two different inputs, radiology reports alone and pairs of order information and radiology reports. Thus, eight experiments were conducted to examine the performance.ResultsWithout order information, BERT achieved the highest area under the precision-recall curve (AUPRC) of 0.5138, which showed a statistically significant improvement over LR, GBDT, and LSTM, and the highest area under the receiver operating characteristic curve (AUROC) of 0.9516. Simply coupling the order information with the radiology reports slightly increased the AUPRC of BERT but did not lead to a statistically significant improvement. This may be due to the complexity of clinical decisions made by radiologists.ConclusionsBERT was assumed to be useful to detect actionable reports. More sophisticated methods are required to use order information effectively.

Loading PDF

Similar Papers
  • Research Article

Intent Recognition and slot filling for ecommerce chatbot

  • May 28, 2024
  • INTERANTIONAL JOURNAL OF SCIENTIFIC RESEARCH IN ENGINEERING AND MANAGEMENT
  • Sapana Bhirud
  • Research Article
  • Citations3

A comparative study of machine learning models for sentiment analysis of transboundary rivers news media articles

  • Dec 01, 2024
  • Soft Computing
  • Jiale Wang +3
  • Research Article
  • Citations13

Do the US president's tweets better predict oil prices? An empirical examination using long short-term memory networks

  • May 30, 2023
  • International Journal of Production Research
  • Stephanie Beyer Díaz +4
  • PDF
  • Research Article
  • Citations11

Extracting Pulmonary Nodules and Nodule Characteristics from Radiology Reports of Lung Cancer Screening Patients Using Transformer Models

  • May 17, 2024
  • Journal of Healthcare Informatics Research
  • Shuang Yang +10
  • PDF
  • Research Article
  • Citations3

An Artificial Neural Network–Based Pediatric Mortality Risk Score: Development and Performance Evaluation Using Data From a Large North American Registry

  • Aug 31, 2021
  • JMIR Medical Informatics
  • Niema Ghanad Poor +4
  • Research Article
  • Citations1

Analisis Perbandingan Model Bert Dan Xlnet Untuk Klasifikasi Tweet Bully Pada Twitter

  • Dec 10, 2024
  • Jurnal Teknologi Informasi dan Ilmu Komputer
  • Teuku Radillah +2
  • Book Chapter

Analyzing Twitter Data for Insights into Public Sentiment During COVID-19 Pandemic

  • Sep 18, 2024
  • Frontiers in artificial intelligence and applications
  • Yang Liu +1
  • Research Article
  • Citations15

Enhancing traditional Chinese medical named entity recognition with Dyn-Att Net: a dynamic attention approach.

  • May 31, 2024
  • PeerJ. Computer science
  • Jingming Hou +2
  • PDF
  • Research Article
  • Citations16

Development and multicenter validation of chest X-ray radiography interpretations based on natural language processing

  • Oct 28, 2021
  • Communications Medicine
  • Yaping Zhang +9
  • PDF
  • Research Article
  • Citations152

Limitations of Transformers on Clinical Text Classification.

  • Feb 26, 2021
  • IEEE Journal of Biomedical and Health Informatics
  • Shang Gao +11
  • Research Article

Disease Risk Prediction Using Structured EHR Data: Can Generalist Large Language Models Match Specialized Clinical Foundation Models? A Comparative Evaluation with Fine-Tuning.

  • May 01, 2026
  • medRxiv : the preprint server for health sciences
  • Bingyu Mao +7
  • PDF
  • Research Article
  • Citations23

Identification of hand-foot syndrome from cancer patients' blog posts: BERT-based deep-learning approach to detect potential adverse drug reaction symptoms.

  • May 04, 2022
  • PLOS ONE
  • Satoshi Nishioka +9
  • Conference Article

Deep Learning Models to Detect Multi-word Expressions

  • Jun 27, 2025
  • Wei Meng +1
  • Research Article

A novel BERT-long short-term memory hybrid model for effective credit card fraud detection

  • Feb 01, 2026
  • IAES International Journal of Artificial Intelligence (IJ-AI)
  • Oussama Ndama +3
  • Book Chapter
  • Citations5

To BERT or Not to BERT Dealing with Possible BERT Failures in an Entailment Task

  • Jan 01, 2020
  • Pedro Fialho +2
Cactus Communications logo

Copyright 2026 Cactus Communications. All rights reserved.