Related Experiment Video
Updated: Jul 31, 2025

Swin-PSAxialNet: An Efficient Multi-Organ Segmentation Technique
Published on: July 5, 2024
DMCT-Net: dual modules convolution transformer network for head and neck tumor segmentation in PET/CT
Jiao Wang1, Yanjun Peng1, Yanfei Guo2
1College of Computer Science and Engineering, Shandong University of Science and Technology, Qingdao 266590 Shandong, People's Republic of China.
Abstract:
Objective.Accurate segmentation of head and neck (H&N) tumors is critical in radiotherapy. However, the existing methods lack effective strategies to integrate local and global information, strong semantic information and context information, and spatial and channel features, which are effective clues to improve the accuracy of tumor segmentation. In this paper, we propose a novel method called dual modules convolution transformer network (DMCT-Net) for H&N tumor segmentation in the fluorodeoxyglucose positron emission tomography/computed tomography (FDG-PET/CT) images.Approach.The DMCT-Net consists of the convolution transformer block (CTB), the squeeze and excitation (SE) pool module, and the multi-attention fusion (MAF) module. First, the CTB is designed to capture the remote dependency and local multi-scale receptive field information by using the standard convolution, the dilated convolution, and the transformer operation. Second, to extract feature information from different angles, we construct the SE pool module, which not only extracts strong semantic features and context features simultaneously but also uses the SE normalization to adaptively fuse features and adjust feature distribution. Third, the MAF module is proposed to combine the global context information, channel information, and voxel-wise local spatial information. Besides, we adopt the up-sampling auxiliary paths to supplement the multi-scale information.Main results.The experimental results show that the method has better or more competitive segmentation performance than several advanced methods on three datasets. The best segmentation metric scores are as follows: DSC of 0.781, HD95 of 3.044, precision of 0.798, and sensitivity of 0.857. Comparative experiments based on bimodal and single modal indicate that bimodal input provides more sufficient and effective information for improving tumor segmentation performance. Ablation experiments verify the effectiveness and significance of each module.Significance.We propose a new network for 3D H&N tumor segmentation in FDG-PET/CT images, which achieves high accuracy.
Insights
This study introduces a novel dual modules convolution transformer network (DMCT-Net) for precise head and neck (H&N) tumor segmentation in FDG-PET/CT scans, significantly improving accuracy over existing methods.
Area of Science:
- Medical Imaging
- Artificial Intelligence in Medicine
- Radiotherapy Planning
Background:
- Accurate segmentation of head and neck (H&N) tumors in FDG-PET/CT images is crucial for effective radiotherapy.
- Existing segmentation methods struggle to integrate local/global information, semantic/contextual features, and spatial/channel data.
Purpose of the Study:
- To propose a novel Dual Modules Convolution Transformer Network (DMCT-Net) for enhanced H&N tumor segmentation.
- To improve the integration of diverse feature types for more accurate tumor delineation.
Main Methods:
- Developed DMCT-Net incorporating Convolution Transformer Blocks (CTB), Squeeze and Excitation (SE) pool module, and Multi-Attention Fusion (MAF) module.
- CTB captures remote dependencies and multi-scale receptive fields; SE pool extracts semantic and contextual features; MAF fuses global context, channel, and spatial information.
- Utilized up-sampling auxiliary paths to supplement multi-scale information.
Main Results:
- DMCT-Net demonstrated superior or competitive segmentation performance against advanced methods on three datasets.
- Achieved high segmentation metric scores: DSC of 0.781, HD95 of 3.044, precision of 0.798, and sensitivity of 0.857.
- Bimodal input (FDG-PET/CT) proved more effective than single modal input; ablation studies confirmed module significance.
Conclusions:
- The proposed DMCT-Net offers a novel and highly accurate approach for 3D H&N tumor segmentation in FDG-PET/CT images.
- The network effectively integrates local, global, semantic, contextual, spatial, and channel features for improved segmentation accuracy.
- This method holds significant potential for advancing radiotherapy planning and treatment for head and neck cancer patients.
More Related Videos
09:21Human Brown Adipose Tissue Depots Automatically Segmented by Positron Emission Tomography/Computed Tomography and Registered Magnetic Resonance Images
Published on: February 18, 2015
04:48Application of Deep Learning-Based Medical Image Segmentation via Orbital Computed Tomography
Published on: November 30, 2022