[1]Ronneberger O, Fischer P, Brox T. U-net: Convolutional networks for biomedical image segmentation[C]//International Conference on Medical image computing and computer-assisted intervention. Cham: Springer international publishing, 2015: 234-241.
[2]刘兆伟,方艳红,郑明宇,等.基于注意力机制与多任务的肺部疾病诊断方法[J].计算机工程,2025,51(01):332342.DOI:10.19678/j.issn.1000-3428.0068786.
LIU Zhaowei, FANG Yanhong, ZHENG Mingyu, et al. Lung Disease Diagnosis Method Based on Attention Mechanism and Multi-Task [J]. Computer Engineering, 2025, 51(01): 332-342. DOI: 10.19678/j.issn.1000-3428.0068786.
[3]林志洁,郑秋岚,梁涌,等.基于内卷U-Net的医学图像分割模型[J].计算机工程,2022,48 (08):180186.DOI:10.19678/j.issn.1000-3428.0062023.
LIN Zhijie, ZHENG Qiulan, LIANG Yong, et al. Medical Image Segmentation Model Based on Involution U-Net [J]. Computer Engineering, 2022, 48(08): 180-186. DOI: 10.19678/j.issn.1000-3428.0062023.
[4]Ibtehaz N, Rahman M S. MultiResUNet: Rethinking the U-Net architecture for multimodal biomedical image segmentation[J]. Neural networks, 2020, 121: 74-87.
[5]Gu Z, Cheng J, Fu H, et al. Ce-net: Context encoder network for 2d medical image segmentation[J]. IEEE transactions on medical imaging, 2019, 38(10): 2281-2292.
[6]Vaswani A, Shazeer N, Parmar N, et al. Attention is all you need[J]. Advances in neural information processing systems, 2017, 30.
[7]陈凌啸,陈小雕.基于Transformer的儿童牙齿全景X线分割算法[J].软件工程,2026,29 (04):49-54.DOI:10.19644/j.cnki.issn2096-1472.2026.004.009.
CHEN Lingxiao, CHEN Xiaodiao. Pediatric Panoramic Dental X-Ray Segmentation Algorithm Based on Transformer [J]. Software Engineering, 2026, 29(04): 49-54. DOI: 10.19644/j.cnki.issn2096-1472.2026.004.009.
[8]Tu Z, Talebi H, Zhang H, et al. Maxvit: Multi-axis vision transformer[C]//European conference on computer vision. Cham: Springer Nature Switzerland, 2022: 459-479.
[9]Huang X, Deng Z, Li D, et al. Missformer: An effective transformer for 2d medical image segmentation[J]. IEEE transactions on medical imaging, 2022, 42(5): 1484-1494.
[10]Chen J, Lu Y, Yu Q, et al. Transunet: Transformers make strong encoders for medical image segmentation[J]. arXiv preprint arXiv:2102.04306, 2021.
[11]Rahman M M, Munir M, Marculescu R. Emcad: Efficient multi-scale convolutional attention decoding for medical image segmentation[C]//Proceedings of the IEEE/CVF conference on computer vision and pattern recognition. 2024: 11769-11779.
[12]Sun G, Pan Y, Kong W, et al. DA-TransUNet: integrating spatial and channel dual attention with transformer U-net for medical image segmentation[J]. Frontiers in Bioengineering and Biotechnology, 2024, 12: 1398237.
[13]Cheng H, Zhu M. FCTrans UNet: A Hybrid CNN and Transformer Model for Medical Image Segmentations[C]//2024 5th International Seminar on Artificial Intelligence, Networking and Information Technology (AINIT). IEEE, 2024: 1277-1282.
[14]Chen J, Liang Z, Lu X. A dual attention and cross layer fusion network with a hybrid CNN and transformer architecture for medical image segmentation[J]. Scientific Reports, 2025, 15(1): 35707.
[15]Li Y, Wan H, Lan L. DCAU-Net: Differential Cross Attention and Channel-Spatial Feature Fusion for Medical Image Segmentation[J]. arXiv preprint arXiv:2603.09530, 2026.
[16]RAHMAN M M, MARCULESCU R. Medical Image Seg mentation via Cascaded Attention Decoding[C]//Pro- ceed ings of the 2023 IEEE/CVF Winter Conference on Appli cations of Computer Vision. Waikoloa, HI, USA: IEEE, 2023: 6211-6220.
[17]Wang W, Xie E, Li X, et al. Pvt v2: Improved baselines with pyramid vision transformer[J]. Computational visual media, 2022, 8(3): 415-424.
[18]Si Y, Xu H, Zhu X, et al. SCSA: Exploring the synergistic effects between spatial and channel attention[J]. Neurocomputing, 2025, 634: 129866.
[19]Liu W, Lu H, Fu H, et al. Learning to upsample by learning to sample[C]//Proceedings of the IEEE/CVF international conference on computer vision. 2023: 6027-6037.
[20]Long J, Shelhamer E, Darrell T. Fully convolutional networks for semantic segmentation[C]//Proceedings of the IEEE conference on computer vision and pattern recognition. 2015: 3431-3440.
[21]Howard A G, Zhu M, Chen B, et al. Mobilenets: Efficient convolutional neural networks for mobile vision applications[J]. arXiv preprint arXiv:1704.04861, 2017.
[22]Zhang X, Zhou X, Lin M, et al. Shufflenet: An extremely efficient convolutional neural network for mobile devices[C]//Proceedings of the IEEE conference on computer vision and pattern recognition. 2018: 6848-6856.
[23]He K, Zhang X, Ren S, et al. Deep residual learning for image recognition[C]//Proceedings of the IEEE conference on computer vision and pattern recognition. 2016: 770-778.
[24]Hu J, Shen L, Sun G. Squeeze-and-excitation networks[C]//Proceedings of the IEEE conference on computer vision and pattern recognition. 2018: 7132-7141.
[25]Shi W, Caballero J, Huszár F, et al. Real-time single image and video super-resolution using an efficient sub-pixel convolutional neural network[C]//Proceedings of the IEEE conference on computer vision and pattern recognition. 2016: 1874-1883.
[26]Rahman M M, Marculescu R. Multi-scale hierarchical vision transformer with cascaded attention decoding for medical image segmentation[C]//Medical Imaging with Deep Learning. PMLR, 2024: 1526-1544.
[27]Landman B, Xu Z, Igelsias J, et al. Multi-atlas labeling beyond the cranial vault-workshop and challenge[C]// Proceedings of the MICCAI Multi-Atlas Labeling Beyond Cranial Vault—Workshop Challenge. 2015: 12.
[28]Bernard O, Lalande A, Zotti C, et al. Deep learning techniques for automatic MRI cardiac multi-structures segmentation and diagnosis: is the problem solved?[J]. IEEE transactions on medical imaging, 2018, 37(11): 2514-2525.
[29]Bernal J, Sánchez J, Vilarino F. Towards automatic polyp detection with a polyp appearance model[J]. Pattern Recognition, 2012, 45(9): 3166-3182.
[30]Russakovsky O, Deng J, Su H, et al. Imagenet large scale visual recognition challenge[J]. International journal of computer vision, 2015, 115(3): 211-252.
[31]OKTAY O, SCHLEMPER J, FOLGOC L L, et al. Attention U-Net: Learning Where to Look for the Pancreas [EB/OL]. [2025-4-10]. https://arxiv.org/pdf/1804. 03999.
[32]AZAD R, HEIDARI M, SHARIATNIA M, et al. Trans DeepLab: Convolution-Free Transformer-Based DeepLab v3+ for Medical Image Segmentation[C]// Proceedings of the 2022 Predictive Intelligence in Medicine. Singapore, Singapore: Springer, 2022: 91-102.
[33]Cao H, Wang Y, Chen J, et al. Swin-unet: Unet-like pure transformer for medical image segmentation[C]//European conference on computer vision. Cham: Springer Nature Switzerland, 2022: 205-218.
[34]Sun H, Xu J, Duan Y. Paratranscnn: Parallelized transcnn encoder for medical image segmentation[J]. arXiv preprint arXiv:2401.15307, 2024.
[35]AZAD R, NIGGEMEIER L, HüTTEMANN M, et al. Be yond Self-Attention: Deformable Large Kernel Attention for Medical Image Segmentation[C]//Pro- ceedings of the 2024 IEEE/CVF Winter Conference on Applications of Computer Vision Waikoloa, HI, USA: IEEE, 2024: 1276 1286.
[36]Shen L, Diao L, Peng R, et al. MCT-Net: a multi-branch hybrid CNN-transformer model for medical image segmentation[J]. Pattern Analysis and Applications, 2025, 28(2): 106.
[37]Tang H, Guo Z, Wang L, et al. Similarity Memory Prior is All You Need for Medical Image Segmentation[C]//Proceedings of the IEEE/CVF International Conference on Computer Vision. 2025: 23009-23018.
[38]Azad R, Jia Y, Aghdam E K, et al. TransCeption: Enhancing medical image segmentation with an inception-like transformer design for efficient feature fusion[J]. Computational Visual Media, 2025.
[39]Alrfou K, Zhao T. GC-UNet: Efficient network for medical image segmentation[J]. Multimedia Tools and Applications, 2026, 85(2): 137.
[40]Alom M Z, Yakopcic C, Hasan M, et al. Recurrent residual U-Net for medical image segmentation[J]. Journal of medical imaging, 2019, 6(1): 014006-014006.
[41]Heidari M, Kazerouni A, Soltany M, et al. Hiformer: Hierarchical multi-scale representations using transformers for medical image segmentation[C]//Proceedings of the IEEE/CVF winter conference on applications of computer vision. 2023: 6202-6212.
[42]Bougourzi F, Dornaika F, Taleb-Ahmed A, et al. Rethinking attention gated with hybrid dual pyramid transformer-cnn for generalized segmentation in medical imaging[C]//International Conference on Pattern Recognition. Cham: Springer Nature Switzerland, 2024: 243-258.
|