LW-CTrans: A lightweight hybrid network of CNN and Transformer for 3D medical image segmentation.

Journal: Medical image analysis

PMID: 40107117

Abstract

Recent models based on convolutional neural network (CNN) and Transformer have achieved the promising performance for 3D medical image segmentation. However, these methods cannot segment small targets well even when equipping large parameters. Therefore, We design a novel lightweight hybrid network that combines the strengths of CNN and Transformers (LW-CTrans) and can boost the global and local representation capability at different stages. Specifically, we first design a dynamic stem that can accommodate images of various resolutions. In the first stage of the hybrid encoder, to capture local features with fewer parameters, we propose a multi-path convolution (MPConv) block. In the middle stages of the hybrid encoder, to learn global and local features meantime, we propose a multi-view pooling based Transformer (MVPFormer) which projects the 3D feature map onto three 2D subspaces to deal with small objects, and use the MPConv block for enhancing local representation learning. In the final stage, to mostly capture global features, only the proposed MVPFormer is used. Finally, to reduce the parameters of the decoder, we propose a multi-stage feature fusion module. Extensive experiments on 3 public datasets for three tasks: stroke lesion segmentation, pancreas cancer segmentation and brain tumor segmentation, show that the proposed LW-CTrans achieves Dices of 62.35±19.51%, 64.69±20.58% and 83.75±15.77% on the 3 datasets, respectively, outperforming 16 state-of-the-art methods, and the numbers of parameters (2.08M, 2.14M and 2.21M on 3 datasets, respectively) are smaller than the non-lightweight 3D methods and close to the lightweight methods. Besides, LW-CTrans also achieves the best performance for small lesion segmentation.

Authors

Hulin Kuang

From the Calgary Stroke Program, Departments of Clinical Neurosciences (W.Q., H.K., E.T., J.M.O., M.G., M.D.H., A.M.D., B.K.M.), Radiology (M.G., M.D.H., A.M.D., B.K.M.), and Community Health Sciences (M.D.H., B.K.M.), University of Calgary, 239 Strathridge Pl SW, Calgary, AB, Canada T3H 4J2; Hotchkiss Brain Institute, Calgary, Alberta, Canada (M.G., M.D.H., A.M.D., B.K.M.), Department of Neurology, Keimyung University, Daegu, South Korea (S.I.S.); and Division of Neuroradiology, Clinic of Radiology and Nuclear Medicine, University Hospital Basel, University of Basel, Basel, Switzerland (J.M.O.).
Yahui Wang

Shanghai Key Laboratory of Forensic Medicine, Shanghai Forensic Service Platform, Academy of Forensic Science, Ministry of Justice, Shanghai, People's Republic of China.
Xianzhen Tan

Hunan Provincial Key Lab on Bioinformatics, School of Computer Science and Engineering, Central South University, Changsha 410000, China.
Jialin Yang

Hunan Provincial Key Lab on Bioinformatics, School of Computer Science and Engineering, Central South University, Changsha 410000, China.
Jiarui Sun

State Key Laboratory of Biogeology and Environmental Geology, School of Earth Sciences, China University of Geosciences, Wuhan, Hubei, China.
Jin Liu

School of Computer Science and Engineering, Central South University, Changsha, China.
Wu Qiu

From the Calgary Stroke Program, Departments of Clinical Neurosciences (W.Q., H.K., E.T., J.M.O., M.G., M.D.H., A.M.D., B.K.M.), Radiology (M.G., M.D.H., A.M.D., B.K.M.), and Community Health Sciences (M.D.H., B.K.M.), University of Calgary, 239 Strathridge Pl SW, Calgary, AB, Canada T3H 4J2; Hotchkiss Brain Institute, Calgary, Alberta, Canada (M.G., M.D.H., A.M.D., B.K.M.), Department of Neurology, Keimyung University, Daegu, South Korea (S.I.S.); and Division of Neuroradiology, Clinic of Radiology and Nuclear Medicine, University Hospital Basel, University of Basel, Basel, Switzerland (J.M.O.).
Jingyang Zhang
Jiulou Zhang
Chunfeng Yang

Laboratory of Image Science and Technology, Southeast University, Nanjing, Jiangsu 210096, P. R. China.
Jianxin Wang
Yang Chen

Orthopedics Department of the First Affiliated Hospital of Tsinghua University, Beijing, China.

Keywords

Algorithms Humans Imaging, Three-Dimensional Neural Networks, Computer

External Resources

View on PubMed Access via DOI PubMed (40107117)

LW-CTrans: A lightweight hybrid network of CNN and Transformer for 3D medical image segmentation.

Abstract

Authors

Keywords

External Resources

Popular Topics

Recent Journals