{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2026,1,19]],"date-time":"2026-01-19T08:24:21Z","timestamp":1768811061684,"version":"3.49.0"},"reference-count":37,"publisher":"MDPI AG","issue":"24","license":[{"start":{"date-parts":[[2020,12,10]],"date-time":"2020-12-10T00:00:00Z","timestamp":1607558400000},"content-version":"vor","delay-in-days":0,"URL":"https:\/\/creativecommons.org\/licenses\/by\/4.0\/"}],"funder":[{"name":"National key technologies R&amp;D program of China","award":["NO.2017YFC0804900"],"award-info":[{"award-number":["NO.2017YFC0804900"]}]},{"DOI":"10.13039\/501100001809","name":"the National Natural Science Foundation of China","doi-asserted-by":"publisher","award":["Project Nos.61872036"],"award-info":[{"award-number":["Project Nos.61872036"]}],"id":[{"id":"10.13039\/501100001809","id-type":"DOI","asserted-by":"publisher"}]}],"content-domain":{"domain":[],"crossmark-restriction":false},"short-container-title":["Sensors"],"abstract":"<jats:p>Due to deep learning\u2019s accurate cognition of the street environment, the convolutional neural network has achieved dramatic development in the application of street scenes. Considering the needs of autonomous driving and assisted driving, in a general way, computer vision technology is used to find obstacles to avoid collisions, which has made semantic segmentation a research priority in recent years. However, semantic segmentation has been constantly facing new challenges for quite a long time. Complex network depth information, large datasets, real-time requirements, etc., are typical problems that need to be solved urgently in the realization of autonomous driving technology. In order to address these problems, we propose an improved lightweight real-time semantic segmentation network, which is based on an efficient image cascading network (ICNet) architecture, using multi-scale branches and a cascaded feature fusion unit to extract rich multi-level features. In this paper, a spatial information network is designed to transmit more prior knowledge of spatial location and edge information. During the course of the training phase, we append an external loss function to enhance the learning process of the deep learning network system as well. This lightweight network can quickly perceive obstacles and detect roads in the drivable area from images to satisfy autonomous driving characteristics. The proposed model shows substantial performance on the Cityscapes dataset. With the premise of ensuring real-time performance, several sets of experimental comparisons illustrate that SP-ICNet enhances the accuracy of road obstacle detection and provides nearly ideal prediction outputs. Compared to the current popular semantic segmentation network, this study also demonstrates the effectiveness of our lightweight network for road obstacle detection in autonomous driving.<\/jats:p>","DOI":"10.3390\/s20247089","type":"journal-article","created":{"date-parts":[[2020,12,10]],"date-time":"2020-12-10T20:18:22Z","timestamp":1607631502000},"page":"7089","update-policy":"https:\/\/doi.org\/10.3390\/mdpi_crossmark_policy","source":"Crossref","is-referenced-by-count":7,"title":["Implementation of a Lightweight Semantic Segmentation Algorithm in Road Obstacle Detection"],"prefix":"10.3390","volume":"20","author":[{"given":"Bushi","family":"Liu","sequence":"first","affiliation":[{"name":"School of Traffic and Transportation, Beijing Jiaotong University, Beijing 100044, China"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Yongbo","family":"Lv","sequence":"additional","affiliation":[{"name":"School of Traffic and Transportation, Beijing Jiaotong University, Beijing 100044, China"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Yang","family":"Gu","sequence":"additional","affiliation":[{"name":"School of Traffic and Transportation, Beijing Jiaotong University, Beijing 100044, China"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"ORCID":"https:\/\/orcid.org\/0000-0002-9272-9215","authenticated-orcid":false,"given":"Wanjun","family":"Lv","sequence":"additional","affiliation":[{"name":"School of Traffic and Transportation, Beijing Jiaotong University, Beijing 100044, China"}],"role":[{"role":"author","vocabulary":"crossref"}]}],"member":"1968","published-online":{"date-parts":[[2020,12,10]]},"reference":[{"key":"ref_1","doi-asserted-by":"crossref","first-page":"5010","DOI":"10.1109\/TIP.2020.2978339","article-title":"RAPNet: Residual Atrous Pyramid Network for Importance-Aware Street Scene Parsing","volume":"29","author":"Zhang","year":"2020","journal-title":"IEEE Trans. Image Process."},{"key":"ref_2","doi-asserted-by":"crossref","unstructured":"He, K.M., Gkioxari, G., Doll\u00e1r, P., and Girshick, R. (2017, January 22\u201329). Mask R-CNN. Proceedings of the IEEE International Conference on Computer Vision (ICCV), Venice, Italy.","DOI":"10.1109\/ICCV.2017.322"},{"key":"ref_3","doi-asserted-by":"crossref","unstructured":"Hariharan, B., Arbel\u00e1ez, P., Girshick, R., and Malik, J. (2014, January 5\u201312). Simultaneous Detection and Segmentation. Proceedings of the European Conference on Computer Vision (ECCV), Zurich, Switzerland.","DOI":"10.1007\/978-3-319-10584-0_20"},{"key":"ref_4","doi-asserted-by":"crossref","first-page":"7431","DOI":"10.1109\/TVT.2019.2926787","article-title":"A Framework for Turning Behavior Classification at Intersections Using 3D LIDAR","volume":"68","author":"Zhang","year":"2019","journal-title":"IEEE Trans. Veh. Technol."},{"key":"ref_5","doi-asserted-by":"crossref","unstructured":"Girshick, R., Donahue, J., Darrell, T., and Malik, J. (2014, January 23\u201328). Rich Feature Hierarchies for Accurate Object Detection and Semantic Segmentation. Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition (CVPR), Columbus, OH, USA.","DOI":"10.1109\/CVPR.2014.81"},{"key":"ref_6","doi-asserted-by":"crossref","first-page":"5128","DOI":"10.1109\/TII.2019.2950031","article-title":"Efficient Outdoor Video Semantic Segmentation Using Feedback-Based Fully Convolution Neural Network","volume":"16","author":"Wong","year":"2020","journal-title":"IEEE Trans. Ind. Inform."},{"key":"ref_7","doi-asserted-by":"crossref","first-page":"36612","DOI":"10.1109\/ACCESS.2020.2968965","article-title":"Multi-Feature View-Based Shallow Convolutional Neural Network for Road Segmentation","volume":"8","author":"Junaid","year":"2020","journal-title":"IEEE Access"},{"key":"ref_8","unstructured":"Tao, C., Qi, J., Li, Y., Wang, H., and Li, H. (August, January 28). Spatial information inference net: Road Extraction Using Road-Specific Contextual Information. Proceedings of the IEEE International Geoscience and Remote Sensing Symposium (IGARSS), Yokohama, Japan."},{"key":"ref_9","doi-asserted-by":"crossref","unstructured":"Liao, J., Cao, L., Li, W., Luo, X., and Feng, X. (2020). UnetDVH-Linear: Linear Feature Segmentation by Dilated Convolution with Vertical and Horizontal Kernels. Sensors, 20.","DOI":"10.3390\/s20205759"},{"key":"ref_10","doi-asserted-by":"crossref","unstructured":"Che, E., Jung, J., and Olsen, M.J. (2019). Object Recognition, Segmentation, and Classification of Mobile Laser Scanning Point Clouds: A State of the Art Review. Sensors, 19.","DOI":"10.3390\/s19040810"},{"key":"ref_11","doi-asserted-by":"crossref","unstructured":"Balado, J., Mart\u00ednez-S\u00e1nchez, J., Arias, P., and Novo, A. (2019). Road Environment Semantic Segmentation with Deep Learning from MLS Point Cloud Data. Sensors, 19.","DOI":"10.3390\/s19163466"},{"key":"ref_12","doi-asserted-by":"crossref","first-page":"2920","DOI":"10.1109\/TGRS.2018.2878510","article-title":"Aerial LaneNet: Lane-Marking Semantic Segmentation in Aerial Imagery Using Wavelet-Enhanced Cost-Sensitive Symmetric Fully Convolutional Neural Networks","volume":"57","author":"Azimi","year":"2019","journal-title":"IEEE Trans. Geosci. Remote Sens."},{"key":"ref_13","doi-asserted-by":"crossref","first-page":"5558","DOI":"10.1109\/LRA.2020.3007457","article-title":"Real-Time Fusion Network for RGB-D Semantic Segmentation Incorporating Unexpected Obstacle Detection for Road-Driving Images","volume":"5","author":"Sun","year":"2020","journal-title":"IEEE Robot. Autom. Lett."},{"key":"ref_14","doi-asserted-by":"crossref","unstructured":"Zhao, H.S., Qi, X.J., Shen, X.Y., Shi, J., and Jia, J. (2017). ICNet for Real-Time Semantic Segmentation on High-Resolution Images. arXiv.","DOI":"10.1007\/978-3-030-01219-9_25"},{"key":"ref_15","doi-asserted-by":"crossref","unstructured":"Chaurasia, A., and Culurciello, E. (2017). LinkNet: Exploiting Encoder Representations for Efficient Semantic Segmentation. arXiv.","DOI":"10.1109\/VCIP.2017.8305148"},{"key":"ref_16","unstructured":"Paszke, A., Chaurasia, A., Kim, S., and Culurciello, E. (2016). ENet: A Deep Neural Network Architecture for Real-Time Semantic Segmentation. arXiv."},{"key":"ref_17","doi-asserted-by":"crossref","unstructured":"Wang, Y., Zhou, Q., Liu, J., Xiong, J., Gao, G., Wu, X., and Latecki, L.J. (2019). LEDNet: A Lightweight Encoder-Decoder Network for Real-Time Semantic Segmentation. arXiv.","DOI":"10.1109\/ICIP.2019.8803154"},{"key":"ref_18","doi-asserted-by":"crossref","unstructured":"Kirillov, A., Wu, Y.X., He, K.M., and Girshick, R. (2020, January 16\u201318). PointRend: Image Segmentation as Rendering. Proceedings of the IEEE\/CVF Conference on Computer Vision and Pattern Recognition (CVPR), Seattle, WA, USA.","DOI":"10.1109\/CVPR42600.2020.00982"},{"key":"ref_19","doi-asserted-by":"crossref","unstructured":"Mostajabi, M., Yadollahpour, P., and Shakhnarovich, G. (2015, January 7\u201312). Feedforward Semantic Segmentation with Zoom-out Features. Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition (CVPR), Boston, MA, USA.","DOI":"10.1109\/CVPR.2015.7298959"},{"key":"ref_20","unstructured":"Simonyan, K., and Zisserman, A. (2014). Very Deep Convolutional Networks for Large-Scale Image Recognition. arXiv."},{"key":"ref_21","doi-asserted-by":"crossref","unstructured":"Arnab, A., Jayasumana, S., Zheng, S., and Torr, P.H. (2016). Higher Order Conditional Random Fields in Deep Neural Networks. arXiv.","DOI":"10.1007\/978-3-319-46475-6_33"},{"key":"ref_22","doi-asserted-by":"crossref","unstructured":"Feng, M.Y., Lu, H.C., and Ding, E.R. (2019, January 15\u201320). Attentive Feedback Network for Boundary-Aware Salient Object Detection. Proceedings of the IEEE\/CVF Conference on Computer Vision and Pattern Recognition (CVPR), Long Beach, CA, USA.","DOI":"10.1109\/CVPR.2019.00172"},{"key":"ref_23","doi-asserted-by":"crossref","unstructured":"Zhao, H.S., Shi, J.P., Qi, X.J., Wang, X., and Jia, J. (2017, January 21\u201326). Pyramid Scene Parsing Network. Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition (CVPR), Honolulu, HI, USA.","DOI":"10.1109\/CVPR.2017.660"},{"key":"ref_24","doi-asserted-by":"crossref","unstructured":"Wang, W.G., Zhao, S.Y., Shen, J.B., Hoi, S.C., and Borji, A. (2019, January 15\u201320). Salient Object Detection with Pyramid Attention and Salient Edges. Proceedings of the IEEE\/CVF Conference on Computer Vision and Pattern Recognition (CVPR), Long Beach, CA, USA.","DOI":"10.1109\/CVPR.2019.00154"},{"key":"ref_25","doi-asserted-by":"crossref","unstructured":"Li, M., Gu, X., Zeng, C., and Feng, Y. (2020). Feasibility Analysis and Application of Reinforcement Learning Algorithm Based on Dynamic Parameter Adjustment. Algorithms, 13.","DOI":"10.3390\/a13090239"},{"key":"ref_26","doi-asserted-by":"crossref","first-page":"105499","DOI":"10.1016\/j.compag.2020.105499","article-title":"Implementation of deep-learning algorithm for obstacle detection and collision avoidance for robotic harvester","volume":"174","author":"Li","year":"2020","journal-title":"Comput. Electron. Agric."},{"key":"ref_27","doi-asserted-by":"crossref","unstructured":"Jampani, V., Sun, D.Q., Liu, M.Y., Yang, M.H., and Kautz, J. (2018, January 8\u201314). Superpixel Sampling Networks. Proceedings of the European Conference on Computer Vision (ECCV), Munich, Germany.","DOI":"10.1007\/978-3-030-01234-2_22"},{"key":"ref_28","doi-asserted-by":"crossref","unstructured":"Liu, Y., Cheng, M.M., Hu, X.W., Wang, K., and Bai, X. (2017, January 21\u201326). Richer Convolutional Features for Edge Detection. Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition (CVPR), Honolulu, HI, USA.","DOI":"10.1109\/CVPR.2017.622"},{"key":"ref_29","doi-asserted-by":"crossref","first-page":"2274","DOI":"10.1109\/TPAMI.2012.120","article-title":"SLIC Superpixels Compared to State-of-the-Art Superpixel Methods","volume":"34","author":"Achanta","year":"2012","journal-title":"IEEE Trans. Pattern Anal. Mach. Intell."},{"key":"ref_30","doi-asserted-by":"crossref","unstructured":"Chen, Z., and Chen, Z.J. (2017, January 14\u201318). RBNet: A Deep Neural Network for Unified Road and Road Boundary Detection. Proceedings of the International Conference on Neural Information Processing (ICONIP), Guangzhou, China.","DOI":"10.1007\/978-3-319-70087-8_70"},{"key":"ref_31","unstructured":"Chen, L., Yang, J., and Kong, H. (July, January 29). Lidar-histogram for fast road and obstacle detection. Proceedings of the IEEE International Conference on Robotics and Automation (ICRA), Singapore."},{"key":"ref_32","doi-asserted-by":"crossref","unstructured":"Chen, Z., Zhang, J., and Tao, D.C. (2019). Progressive LiDAR Adaptation for Road Detection. arXiv.","DOI":"10.1109\/JAS.2019.1911459"},{"key":"ref_33","doi-asserted-by":"crossref","unstructured":"Cordts, M., Omran, M., Ramos, S., Rehfeld, T., Enzweiler, M., Benenson, R., Franke, U., Roth, S., and Schiele, B. (2016). The Cityscapes Dataset for Semantic Urban Scene Understanding. arXiv.","DOI":"10.1109\/CVPR.2016.350"},{"key":"ref_34","doi-asserted-by":"crossref","unstructured":"Albuquerque, A.O., Carvalho J\u00fanior, O.A., Carvalho, O.L.F., de Bem, P.P., Ferreira, P.H.G., de Moura, R.D.S., and Silva, C.R. (2020). Deep Semantic Segmentation of Center Pivot Irrigation Systems from Remotely Sensed Data. Remote Sens., 12.","DOI":"10.3390\/rs12132159"},{"key":"ref_35","doi-asserted-by":"crossref","first-page":"640","DOI":"10.1109\/TPAMI.2016.2572683","article-title":"Fully Convolutional Networks for Semantic Segmentation","volume":"39","author":"Shelhamer","year":"2017","journal-title":"IEEE Trans. Pattern Anal. Mach. Intell."},{"key":"ref_36","unstructured":"Chen, L.C., Papandreou, G., Kokkinos, I., Murphy, K., and Yuille, A.L. (2014). Semantic Image Segmentation with Deep Convolutional Nets and Fully Connected CRFs. arXiv."},{"key":"ref_37","doi-asserted-by":"crossref","first-page":"2481","DOI":"10.1109\/TPAMI.2016.2644615","article-title":"Segnet: A Deep Convolutional Encoder-Decoder Architecture for Image Segmentation","volume":"39","author":"Badrinarayanan","year":"2015","journal-title":"IEEE Trans. Pattern Anal. Mach. Intell."}],"container-title":["Sensors"],"original-title":[],"language":"en","link":[{"URL":"https:\/\/www.mdpi.com\/1424-8220\/20\/24\/7089\/pdf","content-type":"unspecified","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2025,10,11]],"date-time":"2025-10-11T10:43:32Z","timestamp":1760179412000},"score":1,"resource":{"primary":{"URL":"https:\/\/www.mdpi.com\/1424-8220\/20\/24\/7089"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2020,12,10]]},"references-count":37,"journal-issue":{"issue":"24","published-online":{"date-parts":[[2020,12]]}},"alternative-id":["s20247089"],"URL":"https:\/\/doi.org\/10.3390\/s20247089","relation":{},"ISSN":["1424-8220"],"issn-type":[{"value":"1424-8220","type":"electronic"}],"subject":[],"published":{"date-parts":[[2020,12,10]]}}}