Extracting Thyroid Nodules Characteristics from Ultrasound Reports Using Transformer-based Natural Language Processing Methods

03/31/2023
by   Aman Pathak, et al.
0

The ultrasound characteristics of thyroid nodules guide the evaluation of thyroid cancer in patients with thyroid nodules. However, the characteristics of thyroid nodules are often documented in clinical narratives such as ultrasound reports. Previous studies have examined natural language processing (NLP) methods in extracting a limited number of characteristics (<9) using rule-based NLP systems. In this study, a multidisciplinary team of NLP experts and thyroid specialists, identified thyroid nodule characteristics that are important for clinical care, composed annotation guidelines, developed a corpus, and compared 5 state-of-the-art transformer-based NLP methods, including BERT, RoBERTa, LongFormer, DeBERTa, and GatorTron, for extraction of thyroid nodule characteristics from ultrasound reports. Our GatorTron model, a transformer-based large language model trained using over 90 billion words of text, achieved the best strict and lenient F1-score of 0.8851 and 0.9495 for the extraction of a total number of 16 thyroid nodule characteristics, and 0.9321 for linking characteristics to nodules, outperforming other clinical transformer models. To the best of our knowledge, this is the first study to systematically categorize and apply transformer-based NLP models to extract a large number of clinical relevant thyroid nodule characteristics from ultrasound reports. This study lays ground for assessing the documentation quality of thyroid ultrasound reports and examining outcomes of patients with thyroid nodules using electronic health records.

READ FULL TEXT

page 1

page 2

page 3

page 4

research
07/21/2021

An artificial intelligence natural language processing pipeline for information extraction in neuroradiology

The use of electronic health records in medical research is difficult be...
research
09/21/2023

Improving VTE Identification through Adaptive NLP Model Selection and Clinical Expert Rule-based Classifier from Radiology Reports

Rapid and accurate identification of Venous thromboembolism (VTE), a sev...
research
06/15/2018

A Scalable Machine Learning Approach for Inferring Probabilistic US-LI-RADS Categorization

We propose a scalable computerized approach for large-scale inference of...
research
05/14/2019

Extraction and Analysis of Clinically Important Follow-up Recommendations in a Large Radiology Dataset

Communication of follow-up recommendations when abnormalities are identi...
research
12/10/2022

Improving Precancerous Case Characterization via Transformer-based Ensemble Learning

The application of natural language processing (NLP) to cancer pathology...
research
10/27/2022

Working Alliance Transformer for Psychotherapy Dialogue Classification

As a predictive measure of the treatment outcome in psychotherapy, the w...

Please sign up or login with your details

Forgot password? Click here to reset