SGFNet: An Attention-Enhanced Network for Colonic Polyp Segmentation in Endoscopic Images

Bo PANG , Guohua LIU

Journal of Donghua University(English Edition) ›› 2026, Vol. 43 ›› Issue (3) : 101 -112.

PDF (9081KB)
Journal of Donghua University(English Edition) ›› 2026, Vol. 43 ›› Issue (3) :101 -112. DOI: 10.19884/j.1672-5220.202506008
Smart Healthcare
research-article
SGFNet: An Attention-Enhanced Network for Colonic Polyp Segmentation in Endoscopic Images
Author information +
History +
PDF (9081KB)

Abstract

Accurate segmentation of colonic polyps in endoscopic images is vital for early diagnosis and prevention of colorectal cancer(CRC). However, this task remains challenging due to the high variability in polyp morphology, low contrast with surrounding mucosa, and frequent boundary ambiguity. To address these issues, we propose a selective-gated fusion network(SGFNet), a lightweight semantic segmentation network based on the U-Net architecture. SGFNet incorporates two targeted modules: the SSE-Encoder, which integrates selective kernel convolution(SKConv)and squeeze-and-excitation(SE)attention to enhance multi-scale representation and channel recalibration; the PGF-Unit, which employs a gated parallel polarized self-attention(PPSA)mechanism to improve boundary-sensitive feature fusion. The model is evaluated on a unified dataset of 1 612 annotated images from Kvasir-SEG and CVC-ClinicDB. Experimental results demonstrate that SGFNet achieves superior performance in mean intersection over union(mIoU), pixel accuracy(PA), and F1 score, outperforming several representative baseline models. With its balanced design and strong generalization, SGFNet offers a practical and interpretable solution for realworld medical image segmentation and provides insights into attention-based module integration for future computeraided diagnosis systems.

Keywords

colonic polyp segmentation / U-Net / multi-scale feature encoding / attention mechanism / SGFNet / medical image analysis

Cite this article

Download citation ▾
Bo PANG, Guohua LIU. SGFNet: An Attention-Enhanced Network for Colonic Polyp Segmentation in Endoscopic Images. Journal of Donghua University(English Edition), 2026, 43 (3) : 101-112 DOI:10.19884/j.1672-5220.202506008

登录浏览全文

4963

注册一个新账户 忘记密码

References

[1]

Bray F, Laversanne M, Sung H, et al. Global cancer statistics 2022: GLOBOCAN estimates of incidence and mortality worldwide for 36 cancers in 185 countries[J]. CA: A Cancer Journal for Clinicians, 2024, 74(3): 229-263.

[2]

Yang X, Zhu C, Li Y, et al. Incidence and mortality of post-polypectomy colorectal cancer in patients with low-risk adenomas: a systematic review and meta-analysis of observational studies[J]. Digestive Diseases, 2023, 41(2): 206-216.

[3]

Liu J Q, Zhang W W, Liu Y, et al. Polyp segmentation based on implicit edge-guided crosslayer fusion networks[J]. Scientific Reports, 2024, 14: 11678.

[4]

Zhou Z W, Rahman Siddiquee M M, Tajbakhsh N, et al. UNet++: a nested U-Net architecture for medical image segmentation[C]//Deep Learning in Medical Image Analysis and Multimodal Learning for Clinical Decision Support. Cham: Springer, 2018: 3-11.

[5]

Jha D, Smedsrud P H, Riegler M A, et al. ResUNet++: an advanced architecture for medical image segmentation[C]//2019 IEEE International Symposium on Multimedia(ISM). Piscataway, NJ: IEEE, 2020: 225-2255.

[6]

Wei J, Hu Y W, Zhang R M, et al. Shallow attention network for polyp segmentation[C]//Medical Image Computing and Computer Assisted Intervention-MICCAI 2021. Cham: Springer, 2021: 699-708.

[7]

Jha D, Smedsrud P H, Riegler M A, et al. Kvasir-SEG: a segmented polyp dataset[C]//MultiMedia Modeling. Cham: Springer, 2020: 451-462.

[8]

Bernal J, Sánchez F J, Fernández-Esparrach G, et al. WM-DOVA maps for accurate polyp highlighting in colonoscopy: validation vs. saliency maps from physicians[J]. Computerized Medical Imaging and Graphics, 2015, 43: 99-111.

[9]

Woo S, Park J, Lee J Y, et al. CBAM: convolutional block attention module[M]//Computer Vision-ECCV 2018. Cham: Springer International Publishing, 2018: 3-19.

[10]

Song W H, Gao M L, Chehri A, et al. Dualbranch and triple-attention network for pansharpening[J]. Applied Intelligence, 2024, 54(17): 8041-8058.

[11]

Hua X C, Cheng K, Lu H, et al. MSCMNet: multi-scale semantic correlation mining for visible-infrared person re-identification[J]. Pattern Recognition, 2025, 159: 111090.

[12]

Li X, Wang W H, Hu X L, et al. Selective kernel networks[C]//2019 IEEE/CVF Conference on Computer Vision and Pattern Recognition(CVPR). Piscataway, NJ: IEEE, 2020: 510-519.

[13]

Hu J, Shen L, Sun G. Squeeze-and-excitation networks[C]//2018 IEEE/CVF Conference on Computer Vision and Pattern Recognition. Piscataway, NJ: IEEE, 2018: 7132-7141.

[14]

Liu H J, Liu F Q, Fan X Y, et al. Polarized self-attention: towards high-quality pixel-wise regression[PP/OL]. arXiv(2021-07-08)[2025-04-02].

[15]

Dai L, Wang J L, Chen Z Y, et al. Nonequivalent point cloud segmentation method based on convolutional neural network[J]. Journal of Donghua University(Natural Science), 2019, 45(6): 862-868. (in Chinese).

[16]

Rahman M A, Wang Y. Optimizing intersectionover-union in deep neural networks for image segmentation[C]//Advances in Visual Computing. Cham: Springer, 2016: 234-244.

PDF (9081KB)

235

Accesses

0

Citation

Detail

Sections
Recommended

/