TransUNet: Transformers Make Strong Encoders for Medical Image Segmentation

02/10/2021
by   zhouyuyiner, et al.
0

Medical image segmentation is an essential prerequisite for developing healthcare systems, especially for disease diagnosis and treatment planning. On various medical image segmentation tasks, the u-shaped architecture, also known as U-Net, has become the de-facto standard and achieved tremendous success. However, due to the intrinsic locality of convolution operations, U-Net generally demonstrates limitations in explicitly modeling long-range dependency. Transformers, designed for sequence-to-sequence prediction, have emerged as alternative architectures with innate global self-attention mechanisms, but can result in limited localization abilities due to insufficient low-level details. In this paper, we propose TransUNet, which merits both Transformers and U-Net, as a strong alternative for medical image segmentation. On one hand, the Transformer encodes tokenized image patches from a convolution neural network (CNN) feature map as the input sequence for extracting global contexts. On the other hand, the decoder upsamples the encoded features which are then combined with the high-resolution CNN feature maps to enable precise localization. We argue that Transformers can serve as strong encoders for medical image segmentation tasks, with the combination of U-Net to enhance finer details by recovering localized spatial information. TransUNet achieves superior performances to various competing methods on different medical applications including multi-organ segmentation and cardiac segmentation. Code and models are available at https://github.com/Beckschen/TransUNet.

READ FULL TEXT
research
07/19/2021

LeViT-UNet: Make Faster Encoders with Transformer for Medical Image Segmentation

Medical image segmentation plays an essential role in developing compute...
research
02/16/2021

TransFuse: Fusing Transformers and CNNs for Medical Image Segmentation

U-Net based convolutional neural networks with deep feature representati...
research
05/20/2021

Medical Image Segmentation using Squeeze-and-Expansion Transformers

Medical image segmentation is important for computer-aided diagnosis. Go...
research
11/09/2022

RadFormer: Transformers with Global-Local Attention for Interpretable and Accurate Gallbladder Cancer Detection

We propose a novel deep neural network architecture to learn interpretab...
research
09/15/2022

Medical Image Segmentation using LeViT-UNet++: A Case Study on GI Tract Data

Gastro-Intestinal Tract cancer is considered a fatal malignant condition...
research
08/07/2023

Improving FHB Screening in Wheat Breeding Using an Efficient Transformer Model

Fusarium head blight is a devastating disease that causes significant ec...
research
10/27/2022

UNet-2022: Exploring Dynamics in Non-isomorphic Architecture

Recent medical image segmentation models are mostly hybrid, which integr...

Please sign up or login with your details

Forgot password? Click here to reset