Improving Precancerous Case Characterization via Transformer-based Ensemble Learning

12/10/2022
by   Yizhen Zhong, et al.
0

The application of natural language processing (NLP) to cancer pathology reports has been focused on detecting cancer cases, largely ignoring precancerous cases. Improving the characterization of precancerous adenomas assists in developing diagnostic tests for early cancer detection and prevention, especially for colorectal cancer (CRC). Here we developed transformer-based deep neural network NLP models to perform the CRC phenotyping, with the goal of extracting precancerous lesion attributes and distinguishing cancer and precancerous cases. We achieved 0.914 macro-F1 scores for classifying patients into negative, non-advanced adenoma, advanced adenoma and CRC. We further improved the performance to 0.923 using an ensemble of classifiers for cancer status classification and lesion size named entity recognition (NER). Our results demonstrated the potential of using NLP to leverage real-world health record data to facilitate the development of diagnostic tests for early cancer prevention.

READ FULL TEXT

page 1

page 2

page 3

page 4

research
03/06/2023

GlobalNER: Incorporating Non-local Information into Named Entity Recognition

Nowadays, many Natural Language Processing (NLP) tasks see the demand fo...
research
11/10/2019

TENER: Adapting Transformer Encoder for Named Entity Recognition

The Bidirectional long short-term memory networks (BiLSTM) have been wid...
research
03/31/2023

Extracting Thyroid Nodules Characteristics from Ultrasound Reports Using Transformer-based Natural Language Processing Methods

The ultrasound characteristics of thyroid nodules guide the evaluation o...
research
03/01/2017

Skin cancer reorganization and classification with deep neural network

As one kind of skin cancer, melanoma is very dangerous. Dermoscopy based...
research
05/25/2022

A Comparative Study of Gastric Histopathology Sub-size Image Classification: from Linear Regression to Visual Transformer

Gastric cancer is the fifth most common cancer in the world. At the same...
research
07/04/2022

Efficient Lung Cancer Image Classification and Segmentation Algorithm Based on Improved Swin Transformer

With the development of computer technology, various models have emerged...
research
05/30/2023

Machine Learning Approach for Cancer Entities Association and Classification

According to the World Health Organization (WHO), cancer is the second l...

Please sign up or login with your details

Forgot password? Click here to reset