{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2026,5,8]],"date-time":"2026-05-08T17:03:47Z","timestamp":1778259827856,"version":"3.51.4"},"reference-count":52,"publisher":"Oxford University Press (OUP)","issue":"2","license":[{"start":{"date-parts":[[2022,4,7]],"date-time":"2022-04-07T00:00:00Z","timestamp":1649289600000},"content-version":"vor","delay-in-days":6,"URL":"https:\/\/creativecommons.org\/licenses\/by-nc\/4.0\/"}],"funder":[{"DOI":"10.13039\/501100003725","name":"National Research Foundation of Korea","doi-asserted-by":"publisher","award":["2019R1I1A3A01059082"],"award-info":[{"award-number":["2019R1I1A3A01059082"]}],"id":[{"id":"10.13039\/501100003725","id-type":"DOI","asserted-by":"publisher"}]},{"DOI":"10.13039\/501100003710","name":"Korea Health Industry Development Institute","doi-asserted-by":"publisher","award":["HI19C0642"],"award-info":[{"award-number":["HI19C0642"]}],"id":[{"id":"10.13039\/501100003710","id-type":"DOI","asserted-by":"publisher"}]}],"content-domain":{"domain":[],"crossmark-restriction":false},"short-container-title":[],"published-print":{"date-parts":[[2022,4,7]]},"abstract":"<jats:title>Abstract<\/jats:title>\n               <jats:p>Prevention of colorectal cancer (CRC) by inspecting and removing colorectal polyps has become a global health priority because CRC is one of the most frequent cancers in the world. Although recent U-Net-based convolutional neural networks (CNNs) with deep feature representation and skip connections have shown to segment polyps effectively, U-Net-based approaches still have limitations in modeling explicit global contexts, due to the intrinsic nature locality of convolutional operations. To overcome these problems, this study proposes a novel deep learning model, SwinE-Net, for polyp segmentation that effectively combines a CNN-based EfficientNet and Vision Transformer (ViT)-based Swin Ttransformer. The main challenge is to conduct accurate and robust medical segmentation in maintaining global semantics without sacrificing low-level features of CNNs through Swin Transformer. First, the multidilation convolutional block generates refined feature maps to enhance feature discriminability for multilevel feature maps extracted from CNN and ViT. Then, the multifeature aggregation block creates intermediate side outputs from the refined polyp features for efficient training. Finally, the attentive deconvolutional network-based decoder upsamples the refined and combined feature maps to accurately segment colorectal polyps. We compared the proposed approach with previous state-of-the-art methods by evaluating various metrics using five public datasets (Kvasir, ClinicDB, ColonDB, ETIS, and EndoScene). The comparative evaluation, in particular, proved that the proposed approach showed much better performance in the unseen dataset, which shows the generalization and scalability in conducting polyp segmentation. Furthermore, an ablation study was performed to prove the novelty and advantage of the proposed network. The proposed approach outperformed previous studies.<\/jats:p>","DOI":"10.1093\/jcde\/qwac018","type":"journal-article","created":{"date-parts":[[2022,2,3]],"date-time":"2022-02-03T20:11:09Z","timestamp":1643919069000},"page":"616-632","source":"Crossref","is-referenced-by-count":62,"title":["SwinE-Net: hybrid deep learning approach to novel polyp segmentation using convolutional neural network and Swin Transformer"],"prefix":"10.1093","volume":"9","author":[{"given":"Kyeong-Beom","family":"Park","sequence":"first","affiliation":[{"name":"Department of Industrial Engineering, Chonnam National University, 77, Yongbong-ro, Buk-gu, Gwangju 61186, South Korea"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Jae Yeol","family":"Lee","sequence":"additional","affiliation":[{"name":"Department of Industrial Engineering, Chonnam National University, 77, Yongbong-ro, Buk-gu, Gwangju 61186, South Korea"}],"role":[{"role":"author","vocabulary":"crossref"}]}],"member":"286","published-online":{"date-parts":[[2022,4,7]]},"reference":[{"key":"2022040716413586000_bib1","doi-asserted-by":"crossref","first-page":"99","DOI":"10.1016\/j.compmedimag.2015.02.007","article-title":"WM-DOVA maps for accurate polyp highlighting in colonoscopy: Validation vs. saliency maps from physicians","volume":"43","author":"Bernal","year":"2015","journal-title":"Computerized Medical Imaging and Graphics"},{"key":"2022040716413586000_bib2","article-title":"Fully convolutional neural networks for polyp segmentation in colonoscopy","volume-title":"Proceedings of the SPIE Medical Imaging","author":"Brandao","year":"2017"},{"key":"2022040716413586000_bib3","article-title":"Swin-Unet: Unet-like pure transformer for medical image segmentation","author":"Cao","year":"2021"},{"key":"2022040716413586000_bib4","article-title":"TransUNet: Transformers make strong encoders for medical image segmentation","author":"Chen","year":"2021"},{"key":"2022040716413586000_bib5","author":"COVID-19 CT Segmentation Dataset","year":"2020"},{"key":"2022040716413586000_bib6","article-title":"An image is worth 16\u00d716 words: Transformers for image recognition at scale","author":"Dosovitskiy","year":"2020"},{"key":"2022040716413586000_bib7","first-page":"4548","article-title":"Structure-measure: A new way to evaluate foreground maps","volume-title":"Proceedings of the IEEE International Conference on Computer Vision (ICCV)","author":"Fan","year":"2017"},{"key":"2022040716413586000_bib8","doi-asserted-by":"crossref","DOI":"10.24963\/ijcai.2018\/97","article-title":"Enhanced-alignment measure for binary foreground map evaluation","author":"Fan","year":"2018"},{"key":"2022040716413586000_bib9","first-page":"263","article-title":"Pranet: Parallel reverse attention network for polyp segmentation","volume-title":"Proceedings of the International Conference on Medical Image Computing and Computer-Assisted Intervention (MICCAI)","author":"Fan","year":"2020"},{"issue":"8","key":"2022040716413586000_bib10","doi-asserted-by":"crossref","first-page":"2626","DOI":"10.1109\/TMI.2020.2996645","article-title":"Inf-Net: Automatic COVID-19 lung infection segmentation from CT images","volume":"39","author":"Fan","year":"2020","journal-title":"IEEE Transactions on Medical Imaging"},{"key":"2022040716413586000_bib11","doi-asserted-by":"crossref","first-page":"179656","DOI":"10.1109\/ACCESS.2020.3025372","article-title":"MA-Net: A multi-scale attention network for liver and tumor segmentation","volume":"8","author":"Fan","year":"2020","journal-title":"IEEE Access"},{"key":"2022040716413586000_bib12","volume-title":"GLOBOCAN 2008 v1. 2, Cancer incidence and mortality world-wide: IARC cancer base no. 10","author":"Ferlay","year":"2010"},{"issue":"7","key":"2022040716413586000_bib13","doi-asserted-by":"crossref","first-page":"69","DOI":"10.3390\/jimaging6070069","article-title":"Polyp segmentation with fully convolutional deep neural networks\u2014extended evaluation study","volume":"6","author":"Guo","year":"2020","journal-title":"Journal of Imaging"},{"issue":"5","key":"2022040716413586000_bib14","doi-asserted-by":"crossref","first-page":"799","DOI":"10.1136\/gutjnl-2019-319914","article-title":"New artificial intelligence system: First validation study versus experienced endoscopists for colorectal polyp detection","volume":"69","author":"Hassan","year":"2020","journal-title":"Gut"},{"key":"2022040716413586000_bib15","first-page":"770","article-title":"Deep residual learning for image recognition","volume-title":"Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition (CVPR)","author":"He","year":"2016"},{"key":"2022040716413586000_bib17","first-page":"4700","article-title":"Densely connected convolutional networks","volume-title":"Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition (CVPR)","author":"Huang","year":"2017"},{"key":"2022040716413586000_bib16","article-title":"HarDNet-MSEG: A simple encoder\u2013decoder polyp segmentation neural network that achieves over 0.9 mean dice and 86 fps","author":"Huang","year":"2021"},{"key":"2022040716413586000_bib19","doi-asserted-by":"crossref","first-page":"225","DOI":"10.1109\/ISM46123.2019.00049","article-title":"ResUNet++: An advanced architecture for medical image segmentation","volume-title":"Proceedings of the 2019 IEEE International Symposium on Multimedia (ISM)","author":"Jha","year":"2019"},{"key":"2022040716413586000_bib18","doi-asserted-by":"crossref","first-page":"451","DOI":"10.1007\/978-3-030-37734-2_37","article-title":"Kvasir-SEG: A segmented polyp dataset","volume-title":"International Conference on Multimedia Modeling (MMM)","author":"Jha","year":"2020"},{"key":"2022040716413586000_bib20","first-page":"385","article-title":"Receptive field block net for accurate and fast object detection","volume-title":"Proceedings of the European Conference on Computer Vision (ECCV)","author":"Liu","year":"2018"},{"key":"2022040716413586000_bib21","doi-asserted-by":"crossref","DOI":"10.1109\/ICCV48922.2021.00986","article-title":"Swin Transformer: Hierarchical Vision Transformer using shifted windows","author":"Liu","year":"2021"},{"key":"2022040716413586000_bib22","first-page":"3431","article-title":"Fully convolutional networks for semantic segmentation","volume-title":"Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition (CVPR)","author":"Long","year":"2015"},{"key":"2022040716413586000_bib23","doi-asserted-by":"crossref","first-page":"104119","DOI":"10.1016\/j.compbiomed.2020.104119","article-title":"PolypSegNet: A modified encoder\u2013decoder architecture for automated polyp segmentation from colonoscopy images","volume":"128","author":"Mahmud","year":"2021","journal-title":"Computers in Biology and Medicine"},{"key":"2022040716413586000_bib24","article-title":"Transformer transforms salient object detection and camouflaged object detection","author":"Mao","year":"2021"},{"key":"2022040716413586000_bib25","first-page":"248","article-title":"How to evaluate foreground maps?","volume-title":"Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition (CVPR)","author":"Margolin","year":"2014"},{"issue":"10","key":"2022040716413586000_bib26","doi-asserted-by":"crossref","first-page":"713","DOI":"10.1038\/s41551-018-0308-9","article-title":"Detecting colorectal polyps via machine learning","volume":"2","author":"Mori","year":"2018","journal-title":"Nature Biomedical Engineering"},{"key":"2022040716413586000_bib27","doi-asserted-by":"crossref","first-page":"99495","DOI":"10.1109\/ACCESS.2020.2995630","article-title":"Contour-aware polyp segmentation in colonoscopy images using detailed upsampling encoder\u2013decoder networks","volume":"8","author":"Nguyen","year":"2020","journal-title":"IEEE Access"},{"key":"2022040716413586000_bib28","doi-asserted-by":"crossref","first-page":"146308","DOI":"10.1109\/ACCESS.2020.3015108","article-title":"M-GAN: Retinal blood vessel segmentation by balancing losses through stacked deep fully convolutional networks","volume":"8","author":"Park","year":"2020","journal-title":"IEEE Access"},{"key":"2022040716413586000_bib29","author":"Pytorch","year":"2016"},{"key":"2022040716413586000_bib30","first-page":"234","article-title":"U-Net: Convolutional networks for biomedical image segmentation","volume-title":"Proceedings of the International Conference on Medical Image Computing and Computer-Assisted Intervention (MICCAI)","author":"Ronneberger","year":"2015"},{"issue":"4","key":"2022040716413586000_bib31","doi-asserted-by":"crossref","first-page":"1441","DOI":"10.3390\/s21041441","article-title":"A-DenseUNet: Adaptive densely connected UNet for polyp segmentation in colonoscopy images with atrous convolution","volume":"21","author":"Safarov","year":"2021","journal-title":"Sensors"},{"issue":"5","key":"2022040716413586000_bib32","doi-asserted-by":"crossref","first-page":"1316","DOI":"10.1109\/TMI.2019.2948320","article-title":"Modified U-Net (mU-Net) with incorporation of object-dependent high level features for improved liver and liver-tumor segmentation in CT images","volume":"39","author":"Seo","year":"2019","journal-title":"IEEE Transactions on Medical Imaging"},{"issue":"3","key":"2022040716413586000_bib33","doi-asserted-by":"crossref","first-page":"145","DOI":"10.3322\/caac.21601","article-title":"Colorectal cancer statistics","volume":"70","author":"Siegel","year":"2020","journal-title":"A Cancer Journal for Clinicians"},{"issue":"2","key":"2022040716413586000_bib34","doi-asserted-by":"crossref","first-page":"283","DOI":"10.1007\/s11548-013-0926-3","article-title":"Toward embedded detection of polyps in WCE images for early diagnosis of colorectal cancer","volume":"9","author":"Silva","year":"2014","journal-title":"International Journal of Computer Assisted Radiology and Surgery"},{"key":"2022040716413586000_bib35","first-page":"851","article-title":"Colorectal polyp segmentation by U-net with dilation convolution","volume-title":"Proceedings of the IEEE International Conference on Machine Learning And Applications (ICMLA)","author":"Sun","year":"2019"},{"key":"2022040716413586000_bib36","first-page":"4278","article-title":"Inception-v4, Inception-ResNet and the impact of residual connections on learning","volume-title":"Proceedings of the AAAI Conference on Artificial Intelligence (AAAI)","author":"Szegedy","year":"2017"},{"key":"2022040716413586000_bib37","first-page":"6105","article-title":"EfficientNet: Rethinking model scaling for convolutional neural networks","volume-title":"Proceedings of the International Conference on Machine Learning (ICML)","author":"Tan","year":"2019"},{"issue":"2","key":"2022040716413586000_bib38","doi-asserted-by":"crossref","first-page":"630","DOI":"10.1109\/TMI.2015.2487997","article-title":"Automated polyp detection in colonoscopy videos using shape and context information","volume":"35","author":"Tajbakhsh","year":"2015","journal-title":"IEEE Transactions on Medical Imaging"},{"key":"2022040716413586000_bib39","first-page":"307","article-title":"DDANet: Dual decoder attention network for automatic polyp segmentation","volume-title":"Proceedings of the International Conference on Pattern Recognition (ICCV)","author":"Tomar","year":"2021"},{"key":"2022040716413586000_bib40","article-title":"FANet: A feedback attention network for improved biomedical image segmentation","author":"Tomar","year":"2021"},{"key":"2022040716413586000_bib41","first-page":"10347","article-title":"Training data-efficient image transformers & distillation through attention","volume-title":"Proceedings of the International Conference on Machine Learning (ICML)","author":"Touvron","year":"2021"},{"issue":"4","key":"2022040716413586000_bib42","doi-asserted-by":"crossref","first-page":"1023","DOI":"10.1093\/jcde\/qwab030","article-title":"Intervertebral disc instance segmentation using a multistage optimization mask-RCNN (MOM-RCNN)","volume":"8","author":"Vania","year":"2021","journal-title":"Journal of Computational Design and Engineering"},{"issue":"2","key":"2022040716413586000_bib43","doi-asserted-by":"crossref","first-page":"224","DOI":"10.1016\/j.jcde.2018.05.002","article-title":"Automatic spine segmentation from CT images using convolutional neural network via redundant generation of class labels","volume":"6","author":"Vania","year":"2019","journal-title":"Journal of Computational Design and Engineering"},{"issue":"2","key":"2022040716413586000_bib44","doi-asserted-by":"crossref","first-page":"343","DOI":"10.1111\/j.1572-0241.2006.00390.x","article-title":"Polyp miss rate determined by tandem colonoscopy: A systematic review","volume":"101","author":"Van\u00a0Rijn","year":"2006","journal-title":"The American Journal of Gastroenterology"},{"key":"2022040716413586000_bib45","doi-asserted-by":"crossref","first-page":"4037190","DOI":"10.1155\/2017\/4037190","article-title":"A benchmark for endoluminal scene segmentation of colonoscopy images","volume":"2017","author":"V\u00e1zquez","year":"2017","journal-title":"Journal of Healthcare Engineering"},{"key":"2022040716413586000_bib46","first-page":"12321","article-title":"F\u00b3Net: Fusion, feedback and focus for salient object detection","volume-title":"Proceedings of the AAAI Conference on Artificial Intelligence (AAAI)","author":"Wei","year":"2020"},{"key":"2022040716413586000_bib47","first-page":"3","article-title":"CBAM: Convolutional block attention module","volume-title":"Proceedings of the European Conference on Computer Vision (ECCV)","author":"Woo","year":"2018"},{"key":"2022040716413586000_bib48","first-page":"3907","article-title":"Cascaded partial decoder for fast and accurate salient object detection","volume-title":"Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition (CVPR)","author":"Wu","year":"2019"},{"key":"2022040716413586000_bib49","doi-asserted-by":"crossref","DOI":"10.24963\/ijcai.2021\/165","article-title":"Segmenting transparent object in the wild with transformer","author":"Xie","year":"2021"},{"issue":"5","key":"2022040716413586000_bib51","doi-asserted-by":"crossref","first-page":"749","DOI":"10.1109\/LGRS.2018.2802944","article-title":"Road extraction by deep residual U-Net","volume":"15","author":"Zhang","year":"2018","journal-title":"IEEE Geoscience and Remote Sensing Letters"},{"key":"2022040716413586000_bib50","doi-asserted-by":"crossref","DOI":"10.1007\/978-3-030-87193-2_2","article-title":"TransFuse: Fusing transformers and CNNs for medical image segmentation","author":"Zhang","year":"2021"},{"key":"2022040716413586000_bib52","doi-asserted-by":"crossref","first-page":"3","DOI":"10.1007\/978-3-030-00889-5_1","article-title":"UNet++: A nested U-Net architecture for medical image segmentation","volume-title":"Proceedings of Deep Learning in Medical Image Analysis and Multimodal Learning for Clinical Decision Support (DLMIA)","author":"Zhou","year":"2018"}],"container-title":["Journal of Computational Design and Engineering"],"original-title":[],"language":"en","link":[{"URL":"https:\/\/academic.oup.com\/jcde\/article-pdf\/9\/2\/616\/43298266\/qwac018.pdf","content-type":"application\/pdf","content-version":"vor","intended-application":"syndication"},{"URL":"https:\/\/academic.oup.com\/jcde\/article-pdf\/9\/2\/616\/43298266\/qwac018.pdf","content-type":"unspecified","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2022,4,7]],"date-time":"2022-04-07T16:42:31Z","timestamp":1649349751000},"score":1,"resource":{"primary":{"URL":"https:\/\/academic.oup.com\/jcde\/article\/9\/2\/616\/6564811"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2022,4]]},"references-count":52,"journal-issue":{"issue":"2","published-print":{"date-parts":[[2022,4,7]]}},"URL":"https:\/\/doi.org\/10.1093\/jcde\/qwac018","relation":{},"ISSN":["2288-5048"],"issn-type":[{"value":"2288-5048","type":"electronic"}],"subject":[],"published-other":{"date-parts":[[2022,4]]},"published":{"date-parts":[[2022,4]]}}}