{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2026,5,9]],"date-time":"2026-05-09T17:21:28Z","timestamp":1778347288083,"version":"3.51.4"},"reference-count":35,"publisher":"MDPI AG","issue":"5","license":[{"start":{"date-parts":[[2025,5,11]],"date-time":"2025-05-11T00:00:00Z","timestamp":1746921600000},"content-version":"vor","delay-in-days":0,"URL":"https:\/\/creativecommons.org\/licenses\/by\/4.0\/"}],"funder":[{"DOI":"10.13039\/501100003392","name":"Natural Science Foundation of Fujian Province","doi-asserted-by":"publisher","award":["2023J01078"],"award-info":[{"award-number":["2023J01078"]}],"id":[{"id":"10.13039\/501100003392","id-type":"DOI","asserted-by":"publisher"}]},{"DOI":"10.13039\/501100003392","name":"Natural Science Foundation of Fujian Province","doi-asserted-by":"publisher","award":["KFB23151"],"award-info":[{"award-number":["KFB23151"]}],"id":[{"id":"10.13039\/501100003392","id-type":"DOI","asserted-by":"publisher"}]},{"name":"scientific and technological innovation of Fujian Agriculture and Forestry University","award":["2023J01078"],"award-info":[{"award-number":["2023J01078"]}]},{"name":"scientific and technological innovation of Fujian Agriculture and Forestry University","award":["KFB23151"],"award-info":[{"award-number":["KFB23151"]}]}],"content-domain":{"domain":[],"crossmark-restriction":false},"short-container-title":["Algorithms"],"abstract":"<jats:p>The automatic segmentation technique for colorectal polyps in colonoscopy is considered critical for aiding physicians in real-time lesion identification and minimizing diagnostic errors such as false positives and missed lesions. Despite significant progress in existing research, accurate segmentation of colorectal polyps remains technically challenging due to persistent issues such as low contrast between polyps and mucosa, significant morphological heterogeneity, and susceptibility to imaging artifacts caused by bubbles in the colorectal lumen and poor lighting conditions. To address these limitations, this study proposed a novel pyramid vision transformer-based hierarchical path aggregation network (HPANet) for polyp segmentation. Specifically, firstly, the backward multi-scale feature fusion module (BMFM) was developed to enhance the ability of processing polyps with different scales. Secondly, the forward noise reduction module (FNRM) was designed to learn the texture features of the upper and lower layers to reduce the influence of noise such as bubbles. Finally, in order to solve the problem of boundary ambiguity caused by repeated up and down sampling, the boundary feature refinement module (BFRM) was developed to further refine the boundary. The proposed network was compared with several representative networks on five public polyp datasets. Experimental results show that the proposed network achieves better segmentation performance, especially on the Kvasir SEG dataset, where the mDice and mIoU coefficients reach 0.9204 and 0.8655.<\/jats:p>","DOI":"10.3390\/a18050281","type":"journal-article","created":{"date-parts":[[2025,5,12]],"date-time":"2025-05-12T06:17:08Z","timestamp":1747030628000},"page":"281","update-policy":"https:\/\/doi.org\/10.3390\/mdpi_crossmark_policy","source":"Crossref","is-referenced-by-count":3,"title":["HPANet: Hierarchical Path Aggregation Network with Pyramid Vision Transformers for Colorectal Polyp Segmentation"],"prefix":"10.3390","volume":"18","author":[{"given":"Yuhong","family":"Ying","sequence":"first","affiliation":[{"name":"College of Computer and Information Science, Fujian Agriculture and Forestry University, Fuzhou 350002, China"},{"name":"Key Laboratory of Smart Agriculture and Forestry, Fujian Agriculture and Forestry University, Fuzhou 350002, China"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Haoyuan","family":"Li","sequence":"additional","affiliation":[{"name":"College of Computer and Information Science, Fujian Agriculture and Forestry University, Fuzhou 350002, China"},{"name":"Key Laboratory of Smart Agriculture and Forestry, Fujian Agriculture and Forestry University, Fuzhou 350002, China"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"ORCID":"https:\/\/orcid.org\/0000-0002-9132-913X","authenticated-orcid":false,"given":"Yiwen","family":"Zhong","sequence":"additional","affiliation":[{"name":"College of Computer and Information Science, Fujian Agriculture and Forestry University, Fuzhou 350002, China"},{"name":"Key Laboratory of Smart Agriculture and Forestry, Fujian Agriculture and Forestry University, Fuzhou 350002, China"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"ORCID":"https:\/\/orcid.org\/0000-0002-4368-0929","authenticated-orcid":false,"given":"Min","family":"Lin","sequence":"additional","affiliation":[{"name":"College of Computer and Information Science, Fujian Agriculture and Forestry University, Fuzhou 350002, China"},{"name":"Key Laboratory of Smart Agriculture and Forestry, Fujian Agriculture and Forestry University, Fuzhou 350002, China"}],"role":[{"role":"author","vocabulary":"crossref"}]}],"member":"1968","published-online":{"date-parts":[[2025,5,11]]},"reference":[{"key":"ref_1","doi-asserted-by":"crossref","first-page":"233","DOI":"10.3322\/caac.21772","article-title":"Colorectal cancer statistics, 2023","volume":"73","author":"Siegel","year":"2023","journal-title":"CA Cancer J. Clin."},{"key":"ref_2","doi-asserted-by":"crossref","first-page":"669","DOI":"10.1001\/jama.2021.0106","article-title":"Diagnosis and Treatment of Metastatic Colorectal Cancer: A Review","volume":"325","author":"Biller","year":"2021","journal-title":"JAMA"},{"key":"ref_3","first-page":"234","article-title":"U-Net: Convolutional Networks for Biomedical Image Segmentation","volume":"Volume 9351","author":"Navab","year":"2015","journal-title":"Medical Image Computing and Computer-Assisted Intervention\u2014MICCAI 2015"},{"key":"ref_4","unstructured":"Oktay, O., Schlemper, J., Folgoc, L.L., Lee, M., Heinrich, M., Misawa, K., Mori, K., McDonagh, S., Hammerla, N.Y., and Kainz, B. (2018). Attention U-Net: Learning Where to Look for the Pancreas. arXiv."},{"key":"ref_5","doi-asserted-by":"crossref","first-page":"11799","DOI":"10.1109\/JSEN.2020.3015831","article-title":"ABC-Net: Area-Boundary Constraint Network with Dynamical Feature Selection for Colorectal Polyp Segmentation","volume":"21","author":"Fang","year":"2020","journal-title":"IEEE Sens. J."},{"key":"ref_6","doi-asserted-by":"crossref","first-page":"17803","DOI":"10.3934\/mbe.2023791","article-title":"SegT: Separated Edge-Guidance Transformer Network for Polyp Segmentation","volume":"20","author":"Chen","year":"2023","journal-title":"Math. Biosci. Eng"},{"key":"ref_7","first-page":"263","article-title":"PraNet: Parallel Reverse Attention Network for Polyp Segmentation","volume":"12266","author":"Martel","year":"2020","journal-title":"Medical Image Computing and Computer Assisted Intervention\u2014MICCAI 2020"},{"key":"ref_8","unstructured":"Chen, J., Lu, Y., Yu, Q., Luo, X., Adeli, E., Wang, Y., Lu, L., Yuille, A.L., and Zhou, Y. (2021). TransUNet: Transformers Make Strong Encoders for Medical Image Segmentation. arXiv."},{"key":"ref_9","unstructured":"Zhang, Z., and Zhang, W. (2022). Pyramid Medical Transformer for Medical Image Segmentation. arXiv."},{"key":"ref_10","first-page":"448","article-title":"Med-Former: A Transformer Based Architecture for Medical Image Classification","volume":"15011","author":"Linguraru","year":"2024","journal-title":"Medical Image Computing and Computer Assisted Intervention\u2014MICCAI 2024"},{"key":"ref_11","doi-asserted-by":"crossref","first-page":"415","DOI":"10.1007\/s41095-022-0274-8","article-title":"PVT v2: Improved Baselines with Pyramid Vision Transformer","volume":"8","author":"Wang","year":"2022","journal-title":"Comput. Vis. Media"},{"key":"ref_12","doi-asserted-by":"crossref","first-page":"652","DOI":"10.1109\/TPAMI.2019.2938758","article-title":"Res2Net: A New Multi-Scale Backbone Architecture","volume":"43","author":"Gao","year":"2021","journal-title":"IEEE Trans. Pattern Anal. Mach. Intell."},{"key":"ref_13","first-page":"699","article-title":"Shallow Attention Network for Polyp Segmentation","volume":"12901","author":"Cattin","year":"2021","journal-title":"Medical Image Computing and Computer Assisted Intervention\u2014MICCAI 2021"},{"key":"ref_14","first-page":"151","article-title":"TGANet: Text-Guided Attention for Improved Polyp Segmentation","volume":"13433","author":"Wang","year":"2022","journal-title":"Medical Image Computing and Computer Assisted Intervention\u2014MICCAI 2022"},{"key":"ref_15","doi-asserted-by":"crossref","first-page":"2354","DOI":"10.1007\/s10278-024-01124-8","article-title":"UViT-Seg: An Efficient ViT and U-Net-Based Framework for Accurate Colorectal Polyp Segmentation in Colonoscopy and WCE Images","volume":"37","author":"Oukdach","year":"2024","journal-title":"J. Imaging Inform. Med."},{"key":"ref_16","doi-asserted-by":"crossref","first-page":"e14351","DOI":"10.1002\/acm2.14351","article-title":"Multi-Scale Nested UNet with Transformer for Colorectal Polyp Segmentation","volume":"25","author":"Wang","year":"2024","journal-title":"J. Appl. Clin. Med. Phys."},{"key":"ref_17","doi-asserted-by":"crossref","unstructured":"Zhang, W., Fu, C., Zheng, Y., Zhang, F., Zhao, Y., and Sham, C.-W. (2022). HSNet: A Hybrid Semantic Network for Polyp Segmentation. Comput. Biol. Med., 150.","DOI":"10.1016\/j.compbiomed.2022.106173"},{"key":"ref_18","doi-asserted-by":"crossref","unstructured":"Wang, J., Tian, S., Yu, L., Zhou, Z., Wang, F., and Wang, Y. (2023). HIGF-Net: Hierarchical Information-Guided Fusion Network for Polyp Segmentation Based on Transformer and Convolution Feature Learning. Comput. Biol. Med., 161.","DOI":"10.1016\/j.compbiomed.2023.107038"},{"key":"ref_19","doi-asserted-by":"crossref","first-page":"9150015","DOI":"10.26599\/AIR.2023.9150015","article-title":"Polyp-PVT: Polyp Segmentation with Pyramid Vision Transformers","volume":"2","author":"Dong","year":"2023","journal-title":"CAAI Artif. Intell. Res."},{"key":"ref_20","unstructured":"Dosovitskiy, A., Beyer, L., Kolesnikov, A., Weissenborn, D., Zhai, X., Unterthiner, T., Dehghani, M., Minderer, M., Heigold, G., and Gelly, S. (2020). An Image Is Worth 16x16 Words: Transformers for Image Recognition at Scale. arXiv."},{"key":"ref_21","first-page":"15908","article-title":"Transformer in Transformer","volume":"34","author":"Han","year":"2021","journal-title":"Adv. Neural Inf. Process. Syst."},{"key":"ref_22","doi-asserted-by":"crossref","unstructured":"Touvron, H., Cord, M., Sablayrolles, A., Synnaeve, G., and J\u00e9gou, H. (2021, January 11\u201317). Going Deeper with Image Transformers. Proceedings of the IEEE\/CVF International Conference on Computer Vision (ICCV), Montreal, QC, Canada.","DOI":"10.1109\/ICCV48922.2021.00010"},{"key":"ref_23","unstructured":"Zhou, D., Kang, B., Jin, X., Yang, L., Lian, X., Jiang, Z., Hou, Q., and Feng, J. (2021). DeepViT: Towards Deeper Vision Transformer. arXiv."},{"key":"ref_24","doi-asserted-by":"crossref","unstructured":"Liu, Z., Lin, Y., Cao, Y., Hu, H., Wei, Y., Zhang, Z., Lin, S., and Guo, B. (2021, January 11\u201317). Swin Transformer: Hierarchical Vision Transformer Using Shifted Windows. Proceedings of the IEEE\/CVF International Conference on Computer Vision (ICCV), Montreal, QC, Canada.","DOI":"10.1109\/ICCV48922.2021.00986"},{"key":"ref_25","doi-asserted-by":"crossref","unstructured":"Wang, W., Xie, E., Li, X., Fan, D.-P., Song, K., Liang, D., Lu, T., Luo, P., and Shao, L. (2021, January 11\u201317). Pyramid Vision Transformer: A Versatile Backbone for Dense Prediction without Convolutions. Proceedings of the IEEE\/CVF International Conference on Computer Vision (ICCV), Montreal, QC, Canada.","DOI":"10.1109\/ICCV48922.2021.00061"},{"key":"ref_26","doi-asserted-by":"crossref","unstructured":"Woo, S., Park, J., Lee, J.-Y., and Kweon, I.S. (2018, January 8\u201314). CBAM: Convolutional Block Attention Module. Proceedings of the European Conference on Computer Vision (ECCV), Munich, Germany.","DOI":"10.1007\/978-3-030-01234-2_1"},{"key":"ref_27","doi-asserted-by":"crossref","unstructured":"Bhojanapalli, S., Chakrabarti, A., Glasner, D., Li, D., Unterthiner, T., and Veit, A. (2021, January 11\u201317). Understanding Robustness of Transformers for Image Classification. Proceedings of the IEEE\/CVF International Conference on Computer Vision (ICCV), Montreal, QC, Canada.","DOI":"10.1109\/ICCV48922.2021.01007"},{"key":"ref_28","doi-asserted-by":"crossref","unstructured":"Khanh, T.L.B., Dao, D.-P., Ho, N.-H., Yang, H.-J., Baek, E.-T., Lee, G., Kim, S.-H., and Yoo, S.B. (2020). Enhancing U-Net with Spatial-Channel Attention Gate for Abnormal Tissue Segmentation in Medical Imaging. Appl. Sci., 10.","DOI":"10.3390\/app10175729"},{"key":"ref_29","doi-asserted-by":"crossref","first-page":"99","DOI":"10.1016\/j.compmedimag.2015.02.007","article-title":"WM-DOVA Maps for Accurate Polyp Highlighting in Colonoscopy: Validation vs. Saliency Maps from Physicians","volume":"43","author":"Bernal","year":"2015","journal-title":"Comput. Med. Imaging Graph."},{"key":"ref_30","unstructured":"Ro, Y.M., Cheng, W.-H., Kim, J., Chu, W.T., Cui, P., Choi, J.W., Hu, M.C., and De Neve, W. (2020, January 5\u20138). Kvasir-SEG: A Segmented Polyp Dataset. Proceedings of the MultiMedia Modeling, Daejeon, Republic of Korea."},{"key":"ref_31","first-page":"4037190","article-title":"A Benchmark for Endoluminal Scene Segmentation of Colonoscopy Images","volume":"2017","author":"Bernal","year":"2017","journal-title":"J. Healthc. Eng."},{"key":"ref_32","first-page":"4869","article-title":"Automatic Colorectal Polyp Detection in Colonoscopy Video Frames","volume":"17","author":"Geetha","year":"2016","journal-title":"Asian Pac. J. Cancer Prev."},{"key":"ref_33","doi-asserted-by":"crossref","first-page":"283","DOI":"10.1007\/s11548-013-0926-3","article-title":"Toward Embedded Detection of Polyps in WCE Images for Early Diagnosis of Colorectal Cancer","volume":"9","author":"Silva","year":"2014","journal-title":"Int. J. Comput. Assist. Radiol. Surg."},{"key":"ref_34","doi-asserted-by":"crossref","first-page":"94","DOI":"10.1016\/j.isprsjprs.2020.01.013","article-title":"ResUNet-a: A Deep Learning Framework for Semantic Segmentation of Remotely Sensed Data","volume":"162","author":"Diakogiannis","year":"2020","journal-title":"ISPRS J. Photogramm. Remote Sens."},{"key":"ref_35","doi-asserted-by":"crossref","first-page":"014005","DOI":"10.1117\/1.JMI.10.1.014005","article-title":"CaraNet: Context Axial Reverse Attention Network for Segmentation of Small Medical Objects","volume":"10","author":"Lou","year":"2023","journal-title":"J. Med. Imaging"}],"container-title":["Algorithms"],"original-title":[],"language":"en","link":[{"URL":"https:\/\/www.mdpi.com\/1999-4893\/18\/5\/281\/pdf","content-type":"unspecified","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2025,10,9]],"date-time":"2025-10-09T17:30:57Z","timestamp":1760031057000},"score":1,"resource":{"primary":{"URL":"https:\/\/www.mdpi.com\/1999-4893\/18\/5\/281"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2025,5,11]]},"references-count":35,"journal-issue":{"issue":"5","published-online":{"date-parts":[[2025,5]]}},"alternative-id":["a18050281"],"URL":"https:\/\/doi.org\/10.3390\/a18050281","relation":{},"ISSN":["1999-4893"],"issn-type":[{"value":"1999-4893","type":"electronic"}],"subject":[],"published":{"date-parts":[[2025,5,11]]}}}