{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2025,10,12]],"date-time":"2025-10-12T03:19:14Z","timestamp":1760239154214,"version":"build-2065373602"},"reference-count":51,"publisher":"MDPI AG","issue":"20","license":[{"start":{"date-parts":[[2020,10,11]],"date-time":"2020-10-11T00:00:00Z","timestamp":1602374400000},"content-version":"vor","delay-in-days":0,"URL":"https:\/\/creativecommons.org\/licenses\/by\/4.0\/"}],"funder":[{"name":"Science and Technology Major Special Support Projects of Hunan Province","award":["2018GK4010"],"award-info":[{"award-number":["2018GK4010"]}]}],"content-domain":{"domain":[],"crossmark-restriction":false},"short-container-title":["Sensors"],"abstract":"<jats:p>Linear feature extraction is crucial for special objects in semantic segmentation networks, such as slot marking and lanes. The objects with linear characteristics have global contextual information dependency. It is very difficult to capture the complete information of these objects in semantic segmentation tasks. To improve the linear feature extraction ability of the semantic segmentation network, we propose introducing the dilated convolution with vertical and horizontal kernels (DVH) into the task of feature extraction in semantic segmentation networks. Meanwhile, we figure out the outcome if we put the different vertical and horizontal kernels on different places in the semantic segmentation networks. Our networks are trained on the basis of the SS dataset, the TuSimple lane dataset and the Massachusetts Roads dataset. These datasets consist of slot marking, lanes, and road images. The research results show that our method improves the accuracy of the slot marking segmentation of the SS dataset by 2%. Compared with other state-of-the-art methods, our UnetDVH-Linear (v1) obtains better accuracy on the TuSimple Benchmark Lane Detection Challenge with a value of 97.53%. To prove the generalization of our models, road segmentation experiments were performed on aerial images. Without data argumentation, the segmentation accuracy of our model on the Massachusetts roads dataset is 95.3%. Moreover, our models perform better than other models when training with the same loss function and experimental settings. The experiment result shows that the dilated convolution with vertical and horizontal kernels will enhance the neural network on linear feature extraction.<\/jats:p>","DOI":"10.3390\/s20205759","type":"journal-article","created":{"date-parts":[[2020,10,14]],"date-time":"2020-10-14T21:24:39Z","timestamp":1602710679000},"page":"5759","update-policy":"https:\/\/doi.org\/10.3390\/mdpi_crossmark_policy","source":"Crossref","is-referenced-by-count":5,"title":["UnetDVH-Linear: Linear Feature Segmentation by Dilated Convolution with Vertical and Horizontal Kernels"],"prefix":"10.3390","volume":"20","author":[{"ORCID":"https:\/\/orcid.org\/0000-0001-8422-9264","authenticated-orcid":false,"given":"Jiacai","family":"Liao","sequence":"first","affiliation":[{"name":"State Key Laboratory of Advanced Design and Manufacturing for Vehicle Body, Hunan University, Changsha 410006, China"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Libo","family":"Cao","sequence":"additional","affiliation":[{"name":"State Key Laboratory of Advanced Design and Manufacturing for Vehicle Body, Hunan University, Changsha 410006, China"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"ORCID":"https:\/\/orcid.org\/0000-0002-1075-8056","authenticated-orcid":false,"given":"Wei","family":"Li","sequence":"additional","affiliation":[{"name":"State Key Laboratory of Advanced Design and Manufacturing for Vehicle Body, Hunan University, Changsha 410006, China"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Xiaole","family":"Luo","sequence":"additional","affiliation":[{"name":"State Key Laboratory of Advanced Design and Manufacturing for Vehicle Body, Hunan University, Changsha 410006, China"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Xiexing","family":"Feng","sequence":"additional","affiliation":[{"name":"State Key Laboratory of Advanced Design and Manufacturing for Vehicle Body, Hunan University, Changsha 410006, China"}],"role":[{"role":"author","vocabulary":"crossref"}]}],"member":"1968","published-online":{"date-parts":[[2020,10,11]]},"reference":[{"key":"ref_1","doi-asserted-by":"crossref","first-page":"309","DOI":"10.1007\/s00138-018-0986-z","article-title":"Semantic segmentation-based parking space detection with standalone around view monitoring system","volume":"30","author":"Jang","year":"2019","journal-title":"Mach. Vis. Appl."},{"key":"ref_2","doi-asserted-by":"crossref","unstructured":"Wu, Y., Yang, T., Zhao, J., Guan, L., and Jiang, W. (2018, January 26\u201330). VH-HFCN based Parking Slot and Lane Markings Segmentation on Panoramic Surround View. Proceedings of the 2018 IEEE Intelligent Vehicles Symposium (IV), Changshu, China.","DOI":"10.1109\/IVS.2018.8500553"},{"key":"ref_3","doi-asserted-by":"crossref","unstructured":"Neven, D., De Brabandere, B., Georgoulis, S., Proesmans, M., and Van Gool, L. (2018, January 26\u201330). Towards End-to-End Lane Detection: An Instance Segmentation Approach. Proceedings of the 2018 IEEE Intelligent Vehicles Symposium (IV), Changshu, China.","DOI":"10.1109\/IVS.2018.8500547"},{"key":"ref_4","unstructured":"Zhang, W., and Mahale, T. (2018). End to End Video Segmentation for Driving: Lane Detection For Autonomous Car. arXiv."},{"key":"ref_5","doi-asserted-by":"crossref","unstructured":"Chen, P.R., Lo, S.Y., Hang, H.M., Chan, S.W., and Lin, J.J. (2018, January 19\u201321). Efficient road lane marking detection with deep learning. Proceedings of the 2018 IEEE 23rd International Conference on Digital Signal Processing (DSP), Shanghai, China.","DOI":"10.1109\/ICDSP.2018.8631673"},{"key":"ref_6","doi-asserted-by":"crossref","first-page":"128","DOI":"10.1016\/j.isprsjprs.2015.07.002","article-title":"Road networks as collections of minimum cost paths","volume":"108","author":"Wegner","year":"2015","journal-title":"ISPRS J. Photogramm. Remote. Sens."},{"key":"ref_7","doi-asserted-by":"crossref","unstructured":"Wegner, J.D., Montoya-Zegarra, J.A., and Schindler, K. (2013, January 23\u201328). A higher-order crf model for road network extraction. Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition (CVPR), Portland, OR, USA.","DOI":"10.1109\/CVPR.2013.222"},{"key":"ref_8","doi-asserted-by":"crossref","unstructured":"Zhong, Z., Li, J., Cui, W., and Jiang, H. (2016, January 10\u201315). Fully convolutional networks for building and road extraction: Preliminary results. Proceedings of the 2016 IEEE International Geoscience and Remote Sensing Symposium (IGARSS), Beijing, China.","DOI":"10.1109\/IGARSS.2016.7729406"},{"key":"ref_9","doi-asserted-by":"crossref","first-page":"709","DOI":"10.1109\/LGRS.2017.2672734","article-title":"Road Structure Refined CNN for Road Extraction in Aerial Image","volume":"14","author":"Wei","year":"2017","journal-title":"IEEE Geosci. Remote. Sens. Lett."},{"key":"ref_10","doi-asserted-by":"crossref","unstructured":"Canny, J. (1986). A Computational Approach to Edge Detection. IEEE Trans. Pattern Anal. Mach. Intell., 679\u2013698.","DOI":"10.1109\/TPAMI.1986.4767851"},{"key":"ref_11","first-page":"1578","article-title":"Sobel edge detection algorithm","volume":"2","author":"Gupta","year":"2013","journal-title":"Int. J. Comput. Sci. Manag. Res."},{"key":"ref_12","doi-asserted-by":"crossref","first-page":"635","DOI":"10.1016\/j.patcog.2006.06.004","article-title":"Grey-level hit-or-miss transforms\u2014Part i: Unified theory","volume":"40","author":"Naegel","year":"2007","journal-title":"Pattern Recognit."},{"key":"ref_13","doi-asserted-by":"crossref","first-page":"760","DOI":"10.1016\/j.patrec.2009.02.007","article-title":"A hit-or-miss transform for multivariate images","volume":"30","author":"Aptoula","year":"2009","journal-title":"Pattern Recognit. Lett."},{"key":"ref_14","doi-asserted-by":"crossref","first-page":"2145","DOI":"10.1016\/j.patcog.2009.12.023","article-title":"Analysis of new top-hat transformation and the application for infrared dim small target detection","volume":"43","author":"Bai","year":"2010","journal-title":"Pattern Recognit."},{"key":"ref_15","doi-asserted-by":"crossref","first-page":"130","DOI":"10.1016\/j.matcom.2017.12.011","article-title":"Multi-vehicle detection algorithm through combining Harr and HOG features","volume":"155","author":"Wei","year":"2019","journal-title":"Math. Comput. Simul."},{"key":"ref_16","doi-asserted-by":"crossref","first-page":"87","DOI":"10.1016\/S0734-189X(88)80033-1","article-title":"A survey of the hough transform","volume":"44","author":"Illingworth","year":"1988","journal-title":"Comput. Vision Graph. Image Process."},{"key":"ref_17","doi-asserted-by":"crossref","first-page":"310","DOI":"10.1109\/TIP.2006.887731","article-title":"Accurate Centerline Detection and Line Width Estimation of Thick Lines Using the Radon Transform","volume":"16","author":"Zhang","year":"2007","journal-title":"IEEE Trans. Image Process."},{"key":"ref_18","doi-asserted-by":"crossref","first-page":"1034","DOI":"10.1016\/j.patcog.2005.05.014","article-title":"Extended Hough transform for linear feature detection","volume":"39","author":"Cha","year":"2006","journal-title":"Pattern Recognit."},{"key":"ref_19","doi-asserted-by":"crossref","first-page":"436","DOI":"10.1038\/nature14539","article-title":"Deep learning","volume":"521","author":"LeCun","year":"2015","journal-title":"Nature"},{"key":"ref_20","doi-asserted-by":"crossref","first-page":"960","DOI":"10.1016\/j.dsp.2013.01.004","article-title":"Dictionary learning based sparse coefficients for audio classification with max and average pooling","volume":"23","author":"Zubair","year":"2013","journal-title":"Digit. Signal Process."},{"key":"ref_21","unstructured":"Yu, F., and Koltun, V. (2015). Multi-scale context aggregation by dilated convolutions. arXiv."},{"key":"ref_22","doi-asserted-by":"crossref","unstructured":"Ronneberger, O., Fischer, P., and Brox, T. (2015, January 5\u20139). U-Net: Convolutional Networks for Biomedical Image Segmentation. Proceedings of the Medical Image Computing and Computer-Assisted Intervention\u2014MICCAI 2015, Munich, Germany.","DOI":"10.1007\/978-3-319-24574-4_28"},{"key":"ref_23","doi-asserted-by":"crossref","first-page":"640","DOI":"10.1109\/TPAMI.2016.2572683","article-title":"Fully Convolutional Networks for Semantic Segmentation","volume":"39","author":"Shelhamer","year":"2016","journal-title":"IEEE Trans. Pattern Anal. Mach. Intell."},{"key":"ref_24","doi-asserted-by":"crossref","unstructured":"Garcia-Garcia, A., Orts-Escolano, S., Oprea, S., Villena-Mart\u00ednez, V., and Garcia-Rodriguez, J. (2017). A Review on Deep Learning Techniques Applied to Semantic Segmentation. arXiv.","DOI":"10.1016\/j.asoc.2018.05.018"},{"key":"ref_25","doi-asserted-by":"crossref","unstructured":"He, K., Zhang, X., Ren, S., and Sun, J. (July, January 26). Deep Residual Learning for Image Recognition. Proceedings of the 2016 IEEE Conference on Computer Vision and Pattern Recognition (CVPR), Los Alamitos, CA, USA.","DOI":"10.1109\/CVPR.2016.90"},{"key":"ref_26","doi-asserted-by":"crossref","unstructured":"Szegedy, C., Ioffe, S., Vanhoucke, V., and Alemi, A. (2016). Inception-v4, inception-resnet and the impact of residual connections on learning. arXiv.","DOI":"10.1609\/aaai.v31i1.11231"},{"key":"ref_27","doi-asserted-by":"crossref","unstructured":"Huang, G., Liu, Z., Van Der Maaten, L., and Weinberger, K.Q. (2017, January 21\u201326). Densely Connected Convolutional Networks. Proceedings of the 2017 IEEE Conference on Computer Vision and Pattern Recognition (CVPR), Honolulu, HI, USA.","DOI":"10.1109\/CVPR.2017.243"},{"key":"ref_28","doi-asserted-by":"crossref","first-page":"834","DOI":"10.1109\/TPAMI.2017.2699184","article-title":"DeepLab: Semantic Image Segmentation with Deep Convolutional Nets, Atrous Convolution, and Fully Connected CRFs","volume":"40","author":"Chen","year":"2017","journal-title":"IEEE Trans. Pattern Anal. Mach. Intell."},{"key":"ref_29","unstructured":"Paszke, A., Chaurasia, A., Kim, S., and Culurciello, E. (2016). ENet: A Deep Neural Network Architecture for Real-Time Semantic Segmentation. arXiv."},{"key":"ref_30","doi-asserted-by":"crossref","unstructured":"Eigen, D., and Fergus, R. (2015, January 13\u201316). Predicting Depth, Surface Normals and Semantic Labels with a Common Multi-scale Convolutional Architecture. Proceedings of the IEEE international conference on computer vision, Santiago, Chile.","DOI":"10.1109\/ICCV.2015.304"},{"key":"ref_31","doi-asserted-by":"crossref","unstructured":"Roy, A., and Todorovic, S. (2016, January 8\u201316). A Multi-scale CNN for Affordance Segmentation in RGB Images. Proceedings of the Computer Vision \u2013 ECCV 2016, Amsterdam, The Netherlands.","DOI":"10.1007\/978-3-319-46493-0_12"},{"key":"ref_32","doi-asserted-by":"crossref","unstructured":"Bian, X., Lim, S.N., and Zhou, N. (2016, January 7\u20139). Multiscale fully convolutional network with application to industrial inspection. Proceedings of the 2016 IEEE Winter Conference on Applications of Computer Vision (WACV), Lake Placid, NY, USA.","DOI":"10.1109\/WACV.2016.7477595"},{"key":"ref_33","doi-asserted-by":"crossref","unstructured":"He, J., Deng, Z., Zhou, L., Wang, Y., and Qiao, Y. (2019, January 16\u201320). Adaptive Pyramid Context Network for Semantic Segmentation. Proceedings of the 2019 IEEE\/CVF Conference on Computer Vision and Pattern Recognition (CVPR), Long Beach, CA, USA.","DOI":"10.1109\/CVPR.2019.00770"},{"key":"ref_34","doi-asserted-by":"crossref","unstructured":"Hou, Q., Zhang, L., Cheng, M.-M., and Feng, J. (2020, January 14\u201319). Strip Pooling: Rethinking Spatial Pooling for Scene Parsing. Proceedings of the 2020 IEEE\/CVF Conference on Computer Vision and Pattern Recognition (CVPR), Seattle, USA.","DOI":"10.1109\/CVPR42600.2020.00406"},{"key":"ref_35","doi-asserted-by":"crossref","unstructured":"Lo, S.-Y., Hang, H.-M., Chan, S.-W., and Lin, J.-J. (2019, January 27\u201329). Multi-Class Lane Semantic Segmentation using Efficient Convolutional Networks. Proceedings of the 2019 IEEE 21st International Workshop on Multimedia Signal Processing (MMSP), Kuala Lumpur, Malaysia.","DOI":"10.1109\/MMSP.2019.8901686"},{"key":"ref_36","doi-asserted-by":"crossref","first-page":"20","DOI":"10.1016\/j.cogsys.2018.04.004","article-title":"Semantic segmentation via highly fused convolutional network with multiple soft cost functions","volume":"53","author":"Yang","year":"2019","journal-title":"Cogn. Syst. Res."},{"key":"ref_37","doi-asserted-by":"crossref","unstructured":"Wu, Y., Yang, T., Zhao, J., Guan, L., and Li, J. (2017, January 7\u201310). Fully Combined Convolutional Network with Soft Cost Function for Traffic Scene Parsing. Proceedings of the Intelligent Computing Theories and Application, Liverpool, UK.","DOI":"10.1007\/978-3-319-63309-1_64"},{"key":"ref_38","unstructured":"Pizzati, F., Allodi, M., Barrera, A., and Garc\u00eda, F. (2019). Lane Detection and Classification Using Cascaded CNNs. arXiv."},{"key":"ref_39","unstructured":"Mnih, V. (2020, October 10). Machine Learning for Aerial Image Labeling. Available online: http:\/\/citeseerx.ist.psu.edu\/viewdoc\/download?doi=10.1.1.369.1363&rep=rep1&type=pdf."},{"key":"ref_40","doi-asserted-by":"crossref","unstructured":"Szegedy, C., Vanhoucke, V., Ioffe, S., Shlens, J., and Wojna, Z. (July, January 26). Rethinking the Inception Architecture for Computer Vision. Proceedings of the 2016 IEEE Conference on Computer Vision and Pattern Recognition (CVPR), Las Vegas, Nevada.","DOI":"10.1109\/CVPR.2016.308"},{"key":"ref_41","doi-asserted-by":"crossref","unstructured":"Wu, J., and Chung, A.C.S. (2005). Cross Entropy: A New Solver for Markov Random Field Modeling and Applications to Medical Image Segmentation. Comput. Vis., 229\u2013237.","DOI":"10.1007\/11566465_29"},{"key":"ref_42","doi-asserted-by":"crossref","unstructured":"Soomro, T.A., Afifi, A.J., Gao, J., Hellwich, O., Paul, M., and Zheng, L. (2018, January 26\u201329). Strided U-Net Model: Retinal Vessels Segmentation using Dice Loss. Proceedings of the 2018 Digital Image Computing: Techniques and Applications (DICTA), Palm Springs, CA, USA.","DOI":"10.1109\/DICTA.2018.8615770"},{"key":"ref_43","doi-asserted-by":"crossref","first-page":"3627","DOI":"10.1364\/BOE.8.003627","article-title":"ReLayNet: Retinal layer and fluid segmentation of macular optical coherence tomography using fully convolutional networks","volume":"8","author":"Roy","year":"2017","journal-title":"Biomed. Opt. Express"},{"key":"ref_44","doi-asserted-by":"crossref","unstructured":"Guo, Y., Chen, G., Zhao, P., Zhang, W., Miao, J., Wang, J., and Choe, T.E. (2020). Gen-LaneNet: A Generalized and Scalable Approach for 3D Lane Detection. arXiv.","DOI":"10.1007\/978-3-030-58589-1_40"},{"key":"ref_45","doi-asserted-by":"crossref","unstructured":"Qin, T., Chen, T., Chen, Y., and Su, Q. (2020). AVP-SLAM: Semantic Visual Mapping and Localization for Autonomous Vehicles in the Parking Lot. arXiv.","DOI":"10.1109\/IROS45743.2020.9340939"},{"key":"ref_46","doi-asserted-by":"crossref","unstructured":"Ghafoorian, M., Nugteren, C., Baka, N., Booij, O., and Hofmann, M. (2018, January 8\u201314). EL-GAN: Embedding loss driven generative adversarial networks for lane detection. Proceedings of the European Conference on Computer Vision (ECCV) Workshops, Munich, Germany.","DOI":"10.1007\/978-3-030-11009-3_15"},{"key":"ref_47","unstructured":"Paszke, A., Gross, S., Massa, F., Lerer, A., Bradbury, J., Chanan, G., Killeen, T., Lin, Z., Gimelshein, N., and Antiga, L. (2020, October 10). PyTorch: An Imperative Style, High-Performance Deep Learning Library. Available online: https:\/\/static.bsteiner.info\/papers\/pytorch.pdf."},{"key":"ref_48","unstructured":"Kingma, D.P., and Ba, J. (2014). Adam: A method for stochastic optimization. arXiv."},{"key":"ref_49","doi-asserted-by":"crossref","first-page":"41","DOI":"10.1109\/TVT.2019.2949603","article-title":"Robust Lane Detection From Continuous Driving Scenes Using Deep Neural Networks","volume":"69","author":"Zou","year":"2019","journal-title":"IEEE Trans. Veh. Technol."},{"key":"ref_50","doi-asserted-by":"crossref","unstructured":"Pan, X., Shi, J., Luo, P., Wang, X., and Tang, X. (2017). Spatial as deep: Spatial CNN for traffic scene understanding. arXiv.","DOI":"10.1609\/aaai.v32i1.12301"},{"key":"ref_51","doi-asserted-by":"crossref","unstructured":"Hsu, Y.-C., Xu, Z., Kira, Z., and Huang, J. (2018, January 8\u201313). Learning to Cluster for Proposal-Free Instance Segmentation. Proceedings of the 2018 International Joint Conference on Neural Networks (IJCNN), Acre, Brazil.","DOI":"10.1109\/IJCNN.2018.8489379"}],"container-title":["Sensors"],"original-title":[],"language":"en","link":[{"URL":"https:\/\/www.mdpi.com\/1424-8220\/20\/20\/5759\/pdf","content-type":"unspecified","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2025,10,11]],"date-time":"2025-10-11T10:19:13Z","timestamp":1760177953000},"score":1,"resource":{"primary":{"URL":"https:\/\/www.mdpi.com\/1424-8220\/20\/20\/5759"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2020,10,11]]},"references-count":51,"journal-issue":{"issue":"20","published-online":{"date-parts":[[2020,10]]}},"alternative-id":["s20205759"],"URL":"https:\/\/doi.org\/10.3390\/s20205759","relation":{},"ISSN":["1424-8220"],"issn-type":[{"type":"electronic","value":"1424-8220"}],"subject":[],"published":{"date-parts":[[2020,10,11]]}}}