{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2026,7,10]],"date-time":"2026-07-10T15:21:46Z","timestamp":1783696906563,"version":"3.55.0"},"reference-count":43,"publisher":"MDPI AG","issue":"16","license":[{"start":{"date-parts":[[2024,8,16]],"date-time":"2024-08-16T00:00:00Z","timestamp":1723766400000},"content-version":"vor","delay-in-days":0,"URL":"https:\/\/creativecommons.org\/licenses\/by\/4.0\/"}],"funder":[{"DOI":"10.13039\/501100001809","name":"National Natural Science Foundation of China","doi-asserted-by":"publisher","award":["61602157"],"award-info":[{"award-number":["61602157"]}],"id":[{"id":"10.13039\/501100001809","id-type":"DOI","asserted-by":"publisher"}]},{"DOI":"10.13039\/501100001809","name":"National Natural Science Foundation of China","doi-asserted-by":"publisher","award":["62273292"],"award-info":[{"award-number":["62273292"]}],"id":[{"id":"10.13039\/501100001809","id-type":"DOI","asserted-by":"publisher"}]},{"DOI":"10.13039\/501100001809","name":"National Natural Science Foundation of China","doi-asserted-by":"publisher","award":["202102210167"],"award-info":[{"award-number":["202102210167"]}],"id":[{"id":"10.13039\/501100001809","id-type":"DOI","asserted-by":"publisher"}]},{"DOI":"10.13039\/501100001809","name":"National Natural Science Foundation of China","doi-asserted-by":"publisher","award":["242300420284"],"award-info":[{"award-number":["242300420284"]}],"id":[{"id":"10.13039\/501100001809","id-type":"DOI","asserted-by":"publisher"}]},{"DOI":"10.13039\/501100001809","name":"National Natural Science Foundation of China","doi-asserted-by":"publisher","award":["NSFRF240820"],"award-info":[{"award-number":["NSFRF240820"]}],"id":[{"id":"10.13039\/501100001809","id-type":"DOI","asserted-by":"publisher"}]},{"name":"Henan Science and Technology Planning Program","award":["61602157"],"award-info":[{"award-number":["61602157"]}]},{"name":"Henan Science and Technology Planning Program","award":["62273292"],"award-info":[{"award-number":["62273292"]}]},{"name":"Henan Science and Technology Planning Program","award":["202102210167"],"award-info":[{"award-number":["202102210167"]}]},{"name":"Henan Science and Technology Planning Program","award":["242300420284"],"award-info":[{"award-number":["242300420284"]}]},{"name":"Henan Science and Technology Planning Program","award":["NSFRF240820"],"award-info":[{"award-number":["NSFRF240820"]}]},{"DOI":"10.13039\/501100006407","name":"Natural Science Foundation of Henan Province","doi-asserted-by":"publisher","award":["61602157"],"award-info":[{"award-number":["61602157"]}],"id":[{"id":"10.13039\/501100006407","id-type":"DOI","asserted-by":"publisher"}]},{"DOI":"10.13039\/501100006407","name":"Natural Science Foundation of Henan Province","doi-asserted-by":"publisher","award":["62273292"],"award-info":[{"award-number":["62273292"]}],"id":[{"id":"10.13039\/501100006407","id-type":"DOI","asserted-by":"publisher"}]},{"DOI":"10.13039\/501100006407","name":"Natural Science Foundation of Henan Province","doi-asserted-by":"publisher","award":["202102210167"],"award-info":[{"award-number":["202102210167"]}],"id":[{"id":"10.13039\/501100006407","id-type":"DOI","asserted-by":"publisher"}]},{"DOI":"10.13039\/501100006407","name":"Natural Science Foundation of Henan Province","doi-asserted-by":"publisher","award":["242300420284"],"award-info":[{"award-number":["242300420284"]}],"id":[{"id":"10.13039\/501100006407","id-type":"DOI","asserted-by":"publisher"}]},{"DOI":"10.13039\/501100006407","name":"Natural Science Foundation of Henan Province","doi-asserted-by":"publisher","award":["NSFRF240820"],"award-info":[{"award-number":["NSFRF240820"]}],"id":[{"id":"10.13039\/501100006407","id-type":"DOI","asserted-by":"publisher"}]},{"name":"Fundamental Research Funds for the Universities of Henan Province","award":["61602157"],"award-info":[{"award-number":["61602157"]}]},{"name":"Fundamental Research Funds for the Universities of Henan Province","award":["62273292"],"award-info":[{"award-number":["62273292"]}]},{"name":"Fundamental Research Funds for the Universities of Henan Province","award":["202102210167"],"award-info":[{"award-number":["202102210167"]}]},{"name":"Fundamental Research Funds for the Universities of Henan Province","award":["242300420284"],"award-info":[{"award-number":["242300420284"]}]},{"name":"Fundamental Research Funds for the Universities of Henan Province","award":["NSFRF240820"],"award-info":[{"award-number":["NSFRF240820"]}]}],"content-domain":{"domain":[],"crossmark-restriction":false},"short-container-title":["Sensors"],"abstract":"<jats:p>Deep learning has recently made significant progress in semantic segmentation. However, the current methods face critical challenges. The segmentation process often lacks sufficient contextual information and attention mechanisms, low-level features lack semantic richness, and high-level features suffer from poor resolution. These limitations reduce the model\u2019s ability to accurately understand and process scene details, particularly in complex scenarios, leading to segmentation outputs that may have inaccuracies in boundary delineation, misclassification of regions, and poor handling of small or overlapping objects. To address these challenges, this paper proposes a Semantic Segmentation Network Based on Adaptive Attention and Deep Fusion with the Multi-Scale Dilated Convolutional Pyramid (SDAMNet). Specifically, the Dilated Convolutional Atrous Spatial Pyramid Pooling (DCASPP) module is developed to enhance contextual information in semantic segmentation. Additionally, a Semantic Channel Space Details Module (SCSDM) is devised to improve the extraction of significant features through multi-scale feature fusion and adaptive feature selection, enhancing the model\u2019s perceptual capability for key regions and optimizing semantic understanding and segmentation performance. Furthermore, a Semantic Features Fusion Module (SFFM) is constructed to address the semantic deficiency in low-level features and the low resolution in high-level features. The effectiveness of SDAMNet is demonstrated on two datasets, revealing significant improvements in Mean Intersection over Union (MIOU) by 2.89% and 2.13%, respectively, compared to the Deeplabv3+ network.<\/jats:p>","DOI":"10.3390\/s24165305","type":"journal-article","created":{"date-parts":[[2024,8,16]],"date-time":"2024-08-16T04:29:57Z","timestamp":1723782597000},"page":"5305","update-policy":"https:\/\/doi.org\/10.3390\/mdpi_crossmark_policy","source":"Crossref","is-referenced-by-count":3,"title":["Semantic Segmentation Network Based on Adaptive Attention and Deep Fusion Utilizing a Multi-Scale Dilated Convolutional Pyramid"],"prefix":"10.3390","volume":"24","author":[{"given":"Shan","family":"Zhao","sequence":"first","affiliation":[{"name":"School of Software, Henan Polytechnic University, Jiaozuo 454000, China"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Zihao","family":"Wang","sequence":"additional","affiliation":[{"name":"School of Software, Henan Polytechnic University, Jiaozuo 454000, China"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"ORCID":"https:\/\/orcid.org\/0000-0001-9243-5009","authenticated-orcid":false,"given":"Zhanqiang","family":"Huo","sequence":"additional","affiliation":[{"name":"School of Software, Henan Polytechnic University, Jiaozuo 454000, China"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Fukai","family":"Zhang","sequence":"additional","affiliation":[{"name":"School of Software, Henan Polytechnic University, Jiaozuo 454000, China"}],"role":[{"vocabulary":"crossref","role":"author"}]}],"member":"1968","published-online":{"date-parts":[[2024,8,16]]},"reference":[{"key":"ref_1","doi-asserted-by":"crossref","unstructured":"Wang, Y., Yu, X., Yang, Y., Zhang, X., Zhang, Y., Zhang, L., Feng, R., and Xue, J. (2024). A multi-branched semantic segmentation network based on twisted information sharing pattern for medical images. Comput. Methods Programs Biomed., 243.","DOI":"10.1016\/j.cmpb.2023.107914"},{"key":"ref_2","doi-asserted-by":"crossref","first-page":"116534","DOI":"10.1109\/ACCESS.2023.3325677","article-title":"UAV target detection algorithm based on improved YOLOv8","volume":"11","author":"Wang","year":"2023","journal-title":"IEEE Access"},{"key":"ref_3","doi-asserted-by":"crossref","first-page":"2093","DOI":"10.1109\/TCSI.2024.3349588","article-title":"An energy-efficient, unified CNN accelerator for real-time multi-object semantic segmentation for autonomous vehicle","volume":"71","author":"Jung","year":"2024","journal-title":"IEEE Trans. Circuits Syst. I Regul. Pap."},{"key":"ref_4","first-page":"23","article-title":"A threshold selection method from gray-level histograms","volume":"11","author":"Otsu","year":"1975","journal-title":"Automatica"},{"key":"ref_5","first-page":"259","article-title":"Edge detection techniques for image segmentation","volume":"3","author":"Muthukrishnan","year":"2011","journal-title":"Int. J. Comput. Sci. Inf. Technol."},{"key":"ref_6","doi-asserted-by":"crossref","unstructured":"Kaganami, H.G., and Beiji, Z. (2009, January 12\u201314). Region-based segmentation versus edge detection. Proceedings of the 2009 Fifth International Conference on Intelligent Information Hiding and Multimedia Signal Processing, Kyoto, Japan.","DOI":"10.1109\/IIH-MSP.2009.13"},{"key":"ref_7","unstructured":"Roberts, L.G. (1963). Machine Perception of Three-Dimensional Solids. [Ph.D. Thesis, Massachusetts Institute of Technology]."},{"key":"ref_8","doi-asserted-by":"crossref","first-page":"358","DOI":"10.1109\/4.996","article-title":"Design of an image edge detection filter using the Sobel operator","volume":"23","author":"Kanopoulos","year":"1988","journal-title":"IEEE J. Solid-State Circuits"},{"key":"ref_9","first-page":"15","article-title":"Object enhancement and extraction","volume":"10","author":"Prewitt","year":"1970","journal-title":"Pict. Process. Psychopictorics"},{"key":"ref_10","doi-asserted-by":"crossref","first-page":"1463","DOI":"10.1016\/S0031-3203(98)00163-0","article-title":"On the discrete representation of the Laplacian of Gaussian","volume":"32","author":"Gunn","year":"1999","journal-title":"Pattern Recognit."},{"key":"ref_11","doi-asserted-by":"crossref","first-page":"679","DOI":"10.1109\/TPAMI.1986.4767851","article-title":"A computational approach to edge detection","volume":"8","author":"Canny","year":"1986","journal-title":"IEEE Trans. Pattern Anal. Mach. Intell."},{"key":"ref_12","first-page":"1","article-title":"Imagenet classification with deep convolutional neural networks","volume":"25","author":"Krizhevsky","year":"2012","journal-title":"Adv. Neural Inf. Process. Syst."},{"key":"ref_13","doi-asserted-by":"crossref","unstructured":"Yu, C., Wang, J., Peng, C., Gao, C., Yu, G., and Sang, N. (2018, January 8\u201314). Bisenet: Bilateral segmentation network for real-time semantic segmentation. Proceedings of the European Conference on Computer Vision (ECCV), Munich, Germany.","DOI":"10.1007\/978-3-030-01261-8_20"},{"key":"ref_14","doi-asserted-by":"crossref","first-page":"3051","DOI":"10.1007\/s11263-021-01515-2","article-title":"Bisenet v2: Bilateral network with guided aggregation for real-time semantic segmentation","volume":"129","author":"Yu","year":"2021","journal-title":"Int. J. Comput. Vis."},{"key":"ref_15","doi-asserted-by":"crossref","unstructured":"Fu, J., Liu, J., Tian, H., Li, Y., Bao, Y., Fang, Z., and Lu, H. (2019, January 15\u201320). Dual attention network for scene segmentation. Proceedings of the IEEE\/CVF Conference on Computer Vision and Pattern Recognition, Long Beach, CA, USA.","DOI":"10.1109\/CVPR.2019.00326"},{"key":"ref_16","unstructured":"Ding, X., Guo, Y., Ding, G., and Han, J. (November, January 27). Acnet: Strengthening the kernel skeletons for powerful cnn via asymmetric convolution blocks. Proceedings of the IEEE\/CVF International Conference on Computer Vision, Seoul, Republic of Korea."},{"key":"ref_17","doi-asserted-by":"crossref","unstructured":"Yuan, Y., Chen, X., Chen, X., and Wang, J. (2021). Segmentation transformer: Object-contextual representations for semantic segmentation. arXiv.","DOI":"10.1007\/978-3-030-58539-6_11"},{"key":"ref_18","doi-asserted-by":"crossref","unstructured":"Lyu, Y., Vosselman, G., Xia, G.-S., and Yang, M.Y. (2021). Bidirectional multi-scale attention networks for semantic segmentation of oblique UAV imagery. arXiv.","DOI":"10.5194\/isprs-annals-V-2-2021-75-2021"},{"key":"ref_19","first-page":"9355","article-title":"Twins: Revisiting the design of spatial attention in vision transformers","volume":"34","author":"Chu","year":"2021","journal-title":"Adv. Neural Inf. Process. Syst."},{"key":"ref_20","first-page":"5603018","article-title":"Hybrid multiple attention network for semantic segmentation in aerial images","volume":"60","author":"Niu","year":"2021","journal-title":"IEEE Trans. Geosci. Remote Sens."},{"key":"ref_21","unstructured":"Hong, Y., Pan, H., Sun, W., and Jia, Y. (2021). Deep dual-resolution networks for real-time and accurate semantic segmentation of road scenes. arXiv."},{"key":"ref_22","doi-asserted-by":"crossref","unstructured":"Ronneberger, O., Fischer, P., and Brox, T. (2015, January 5\u20139). U-net: Convolutional networks for biomedical image segmentation. Proceedings of the Medical Image Computing and Computer-Assisted Intervention\u2013MICCAI 2015: 18th International Conference, Munich, Germany. Part III 18.","DOI":"10.1007\/978-3-319-24574-4_28"},{"key":"ref_23","doi-asserted-by":"crossref","first-page":"107404","DOI":"10.1016\/j.patcog.2020.107404","article-title":"U2-Net: Going deeper with nested U-structure for salient object detection","volume":"106","author":"Qin","year":"2020","journal-title":"Pattern Recognit."},{"key":"ref_24","first-page":"12077","article-title":"SegFormer: Simple and efficient design for semantic segmentation with transformers","volume":"34","author":"Xie","year":"2021","journal-title":"Adv. Neural Inf. Process. Syst."},{"key":"ref_25","doi-asserted-by":"crossref","first-page":"4115","DOI":"10.1109\/JSTARS.2021.3069909","article-title":"Uvid-net: Enhanced semantic segmentation of uav aerial videos by embedding temporal information","volume":"14","author":"Girisha","year":"2021","journal-title":"IEEE J. Sel. Top. Appl. Earth Obs. Remote Sens."},{"key":"ref_26","doi-asserted-by":"crossref","unstructured":"Zheng, S., Lu, J., Zhao, H., Zhu, X., Luo, Z., Wang, Y., Fu, Y., Feng, J., Xiang, T., and Torr, P.H. (2021, January 19\u201325). Rethinking semantic segmentation from a sequence-to-sequence perspective with transformers. Proceedings of the IEEE\/CVF Conference on Computer Vision and Pattern Recognition, Nashville, TN, USA.","DOI":"10.1109\/CVPR46437.2021.00681"},{"key":"ref_27","doi-asserted-by":"crossref","unstructured":"Liu, Z., Lin, Y., Cao, Y., Hu, H., Wei, Y., Zhang, Z., Lin, S., and Guo, B. (2021, January 10\u201317). Swin transformer: Hierarchical vision transformer using shifted windows. Proceedings of the IEEE\/CVF International Conference on Computer Vision, Montreal, QC, Canada.","DOI":"10.1109\/ICCV48922.2021.00986"},{"key":"ref_28","doi-asserted-by":"crossref","unstructured":"Liu, Z., Hu, H., Lin, Y., Yao, Z., Xie, Z., Wei, Y., Ning, J., Cao, Y., Zhang, Z., and Dong, L. (2022, January 18\u201324). Swin transformer v2: Scaling up capacity and resolution. Proceedings of the IEEE\/CVF Conference on Computer Vision and Pattern Recognition, New Orleans, LA, USA.","DOI":"10.1109\/CVPR52688.2022.01170"},{"key":"ref_29","doi-asserted-by":"crossref","unstructured":"Cao, H., Wang, Y., Chen, J., Jiang, D., Zhang, X., Tian, Q., and Wang, M. (2022, January 23\u201327). Swin-unet: Unet-like pure transformer for medical image segmentation. Proceedings of the European Conference on Computer Vision; Springer, Tel Aviv, Israel.","DOI":"10.1007\/978-3-031-25066-8_9"},{"key":"ref_30","doi-asserted-by":"crossref","first-page":"2481","DOI":"10.1109\/TPAMI.2016.2644615","article-title":"Segnet: A deep convolutional encoder-decoder architecture for image segmentation","volume":"39","author":"Badrinarayanan","year":"2017","journal-title":"IEEE Trans. Pattern Anal. Mach. Intell."},{"key":"ref_31","first-page":"1140","article-title":"Segnext: Rethinking convolutional attention design for semantic segmentation","volume":"35","author":"Guo","year":"2022","journal-title":"Adv. Neural Inf. Process. Syst."},{"key":"ref_32","doi-asserted-by":"crossref","unstructured":"Chen, L.-C., Zhu, Y., Papandreou, G., Schroff, F., and Adam, H. (2018, January 8\u201314). Encoder-decoder with atrous separable convolution for semantic image segmentation. Proceedings of the European Conference on Computer Vision (ECCV), Munich, Germany.","DOI":"10.1007\/978-3-030-01234-2_49"},{"key":"ref_33","doi-asserted-by":"crossref","unstructured":"Luo, H., and Lu, Y. (2023, January 8\u201310). DeepLabV3-SAM: A novel image segmentation method for rail transportation. Proceedings of the 2023 3rd International Conference on Electronic Information Engineering and Computer Communication (EIECC), Wuhan, China.","DOI":"10.1109\/EIECC60864.2023.10456611"},{"key":"ref_34","doi-asserted-by":"crossref","unstructured":"Hou, X., Chen, P., and Gu, H. (2024). LM-DeeplabV3+: A Lightweight Image Segmentation Algorithm Based on Multi-Scale Feature Interaction. Appl. Sci., 14.","DOI":"10.20944\/preprints202401.1052.v1"},{"key":"ref_35","doi-asserted-by":"crossref","unstructured":"Sandler, M., Howard, A., Zhu, M., Zhmoginov, A., and Chen, L.-C. (2018, January 18\u201323). Mobilenetv2: Inverted residuals and linear bottlenecks. Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition, Salt Lake City, UT, USA.","DOI":"10.1109\/CVPR.2018.00474"},{"key":"ref_36","doi-asserted-by":"crossref","unstructured":"Wang, Q., Wu, B., Zhu, P., Li, P., Zuo, W., and Hu, Q. (2020, January 13\u201319). ECA-Net: Efficient channel attention for deep convolutional neural networks. Proceedings of the IEEE\/CVF Conference on Computer Vision and Pattern Recognition, Seattle, WA, USA.","DOI":"10.1109\/CVPR42600.2020.01155"},{"key":"ref_37","doi-asserted-by":"crossref","unstructured":"Zhang, H., Zu, K., Lu, J., Zou, Y., and Meng, D. (2022, January 4\u20138). EPSANet: An efficient pyramid squeeze attention block on convolutional neural network. Proceedings of the Asian Conference on Computer Vision, Macao, China.","DOI":"10.1007\/978-3-031-26313-2_33"},{"key":"ref_38","doi-asserted-by":"crossref","unstructured":"He, K., Zhang, X., Ren, S., and Sun, J. (2016, January 27\u201330). Deep residual learning for image recognition. Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition, Las Vegas, NV, USA.","DOI":"10.1109\/CVPR.2016.90"},{"key":"ref_39","doi-asserted-by":"crossref","unstructured":"Woo, S., Park, J., Lee, J.-Y., and Kweon, I.S. (2018, January 8\u201314). Cbam: Convolutional block attention module. Proceedings of the European Conference on Computer Vision (ECCV), Munich, Germany.","DOI":"10.1007\/978-3-030-01234-2_1"},{"key":"ref_40","doi-asserted-by":"crossref","unstructured":"Li, X., Wang, W., Hu, X., and Yang, J. (2019, January 15\u201320). Selective kernel networks. Proceedings of the IEEE\/CVF Conference on Computer Vision and Pattern Recognition, Long Beach, CA, USA.","DOI":"10.1109\/CVPR.2019.00060"},{"key":"ref_41","doi-asserted-by":"crossref","first-page":"1536976","DOI":"10.1155\/2022\/1536976","article-title":"Using AAEHS-net as an attention-based auxiliary extraction and hybrid subsampled network for semantic segmentation","volume":"2022","author":"Zhao","year":"2022","journal-title":"Comput. Intell. Neurosci."},{"key":"ref_42","doi-asserted-by":"crossref","unstructured":"Chen, Y., Wang, Y., Lu, P., Chen, Y., and Wang, G. (2018, January 23\u201326). Large-scale structure from motion with semantic constraints of aerial images. Proceedings of the Chinese Conference on Pattern Recognition and Computer Vision (PRCV), Guangzhou, China.","DOI":"10.1007\/978-3-030-03398-9_30"},{"key":"ref_43","unstructured":"Ruder, S. (2016). An overview of gradient descent optimization algorithms. arXiv."}],"container-title":["Sensors"],"original-title":[],"language":"en","link":[{"URL":"https:\/\/www.mdpi.com\/1424-8220\/24\/16\/5305\/pdf","content-type":"unspecified","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2025,10,10]],"date-time":"2025-10-10T15:37:31Z","timestamp":1760110651000},"score":1,"resource":{"primary":{"URL":"https:\/\/www.mdpi.com\/1424-8220\/24\/16\/5305"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2024,8,16]]},"references-count":43,"journal-issue":{"issue":"16","published-online":{"date-parts":[[2024,8]]}},"alternative-id":["s24165305"],"URL":"https:\/\/doi.org\/10.3390\/s24165305","relation":{},"ISSN":["1424-8220"],"issn-type":[{"value":"1424-8220","type":"electronic"}],"subject":[],"published":{"date-parts":[[2024,8,16]]}}}