{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2026,7,16]],"date-time":"2026-07-16T21:14:16Z","timestamp":1784236456513,"version":"3.55.0"},"reference-count":35,"publisher":"MDPI AG","issue":"9","license":[{"start":{"date-parts":[[2019,4,29]],"date-time":"2019-04-29T00:00:00Z","timestamp":1556496000000},"content-version":"vor","delay-in-days":0,"URL":"https:\/\/creativecommons.org\/licenses\/by\/4.0\/"}],"funder":[{"DOI":"10.13039\/501100001809","name":"National Natural Science Foundation of China","doi-asserted-by":"publisher","award":["61673017"],"award-info":[{"award-number":["61673017"]}],"id":[{"id":"10.13039\/501100001809","id-type":"DOI","asserted-by":"publisher"}]},{"DOI":"10.13039\/501100001809","name":"National Natural Science Foundation of China","doi-asserted-by":"publisher","award":["61403398"],"award-info":[{"award-number":["61403398"]}],"id":[{"id":"10.13039\/501100001809","id-type":"DOI","asserted-by":"publisher"}]},{"DOI":"10.13039\/501100007128","name":"Natural Science Foundation of Shaanxi Province","doi-asserted-by":"publisher","award":["2017JM6077"],"award-info":[{"award-number":["2017JM6077"]}],"id":[{"id":"10.13039\/501100007128","id-type":"DOI","asserted-by":"publisher"}]},{"DOI":"10.13039\/501100007128","name":"Natural Science Foundation of Shaanxi Province","doi-asserted-by":"publisher","award":["2018ZDXM-GY-039"],"award-info":[{"award-number":["2018ZDXM-GY-039"]}],"id":[{"id":"10.13039\/501100007128","id-type":"DOI","asserted-by":"publisher"}]}],"content-domain":{"domain":[],"crossmark-restriction":false},"short-container-title":["Remote Sensing"],"abstract":"<jats:p>The technology used for road extraction from remote sensing images plays an important role in urban planning, traffic management, navigation, and other geographic applications. Although deep learning methods have greatly enhanced the development of road extractions in recent years, this technology is still in its infancy. Because the characteristics of road targets are complex, the accuracy of road extractions is still limited. In addition, the ambiguous prediction of semantic segmentation methods also makes the road extraction result blurry. In this study, we improved the performance of the road extraction network by integrating atrous spatial pyramid pooling (ASPP) with an Encoder-Decoder network. The proposed approach takes advantage of ASPP\u2019s ability to extract multiscale features and the Encoder-Decoder network\u2019s ability to extract detailed features. Therefore, it can achieve accurate and detailed road extraction results. For the first time, we utilized the structural similarity (SSIM) as a loss function for road extraction. Therefore, the ambiguous predictions in the extraction results can be removed, and the image quality of the extracted roads can be improved. The experimental results using the Massachusetts Road dataset show that our method achieves an F1-score of 83.5% and an SSIM of 0.893. Compared with the normal U-net, our method improves the F1-score by 2.6% and the SSIM by 0.18. Therefore, it is demonstrated that the proposed approach can extract roads from remote sensing images more effectively and clearly than the other compared methods.<\/jats:p>","DOI":"10.3390\/rs11091015","type":"journal-article","created":{"date-parts":[[2019,4,29]],"date-time":"2019-04-29T07:01:22Z","timestamp":1556521282000},"page":"1015","update-policy":"https:\/\/doi.org\/10.3390\/mdpi_crossmark_policy","source":"Crossref","is-referenced-by-count":161,"title":["Road Extraction by Using Atrous Spatial Pyramid Pooling Integrated Encoder-Decoder Network and Structural Similarity Loss"],"prefix":"10.3390","volume":"11","author":[{"ORCID":"https:\/\/orcid.org\/0000-0003-2051-7843","authenticated-orcid":false,"given":"Hao","family":"He","sequence":"first","affiliation":[{"name":"Department of Control Engineering, Rocket Force University of Engineering, Xi\u2019an 710025, China"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"ORCID":"https:\/\/orcid.org\/0000-0002-4948-4056","authenticated-orcid":false,"given":"Dongfang","family":"Yang","sequence":"additional","affiliation":[{"name":"Department of Control Engineering, Rocket Force University of Engineering, Xi\u2019an 710025, China"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Shicheng","family":"Wang","sequence":"additional","affiliation":[{"name":"Department of Control Engineering, Rocket Force University of Engineering, Xi\u2019an 710025, China"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Shuyang","family":"Wang","sequence":"additional","affiliation":[{"name":"Department of Information Engineering, Rocket Force University of Engineering, Xi\u2019an 710025, China"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Yongfei","family":"Li","sequence":"additional","affiliation":[{"name":"Department of Control Engineering, Rocket Force University of Engineering, Xi\u2019an 710025, China"}],"role":[{"vocabulary":"crossref","role":"author"}]}],"member":"1968","published-online":{"date-parts":[[2019,4,29]]},"reference":[{"key":"ref_1","doi-asserted-by":"crossref","first-page":"113","DOI":"10.1109\/34.659930","article-title":"An unbiased detector of curvilinear structures","volume":"20","author":"Steger","year":"1998","journal-title":"IEEE Trans. Pattern Anal. Mach. Intell."},{"key":"ref_2","doi-asserted-by":"crossref","first-page":"84","DOI":"10.1109\/34.23115","article-title":"Edge detection and linear feature extraction using a 2-D random field model","volume":"11","author":"Zhou","year":"1989","journal-title":"IEEE Trans. Pattern Anal. Mach. Intell."},{"key":"ref_3","first-page":"2","article-title":"Automatic road extraction based on cross detection in suburb","volume":"36","author":"Koutaki","year":"2004","journal-title":"Electron. Imaging"},{"key":"ref_4","doi-asserted-by":"crossref","first-page":"4144","DOI":"10.1109\/TGRS.2007.906107","article-title":"Road network extraction and intersection detection from aerial images by tracking road footprints","volume":"45","author":"Hu","year":"2007","journal-title":"IEEE Trans. Geosci. Remote Sens."},{"key":"ref_5","doi-asserted-by":"crossref","first-page":"405","DOI":"10.1111\/j.1477-9730.2008.00496.x","article-title":"Road junction extraction from high-resolution aerial imagery","volume":"23","author":"Ravanbakhsh","year":"2010","journal-title":"Photogramm. Rec."},{"key":"ref_6","unstructured":"Marikhu, R., Dailey, M.N., Makhanov, S., and Honda, K. (2007, January 18\u201322). A Family of quadratic snakes for road extraction. Proceedings of the 8th Asia Conference on Computer Vision, Tokyo, Japan."},{"key":"ref_7","doi-asserted-by":"crossref","unstructured":"Wegner, J.D., Montoya-Zegarra, J.A., and Schindler, K. (2013, January 23\u201328). A higher-order CRF model for road network extraction. Proceedings of the IEEE Computer Society Conference on Computer Vision and Pattern Recognition, Portland, OR, USA.","DOI":"10.1109\/CVPR.2013.222"},{"key":"ref_8","doi-asserted-by":"crossref","unstructured":"Mnih, V., and Hinton, G.E. (2010, January 5\u201311). Learning to detect roads in high-resolution aerial images. Proceedings of the European Conference on Computer Vision, Heraklion, Crete, Greece.","DOI":"10.1007\/978-3-642-15567-3_16"},{"key":"ref_9","unstructured":"Mnih, V. (2013). Machine Learning for Aerial Image Labeling. [Ph.D. Thesis, University of Toronto]."},{"key":"ref_10","doi-asserted-by":"crossref","first-page":"3144","DOI":"10.1080\/01431161.2015.1054049","article-title":"Road network extraction: A neural-dynamic framework based on deep learning and a finite state machine","volume":"36","author":"Wang","year":"2015","journal-title":"Int. J. Remote Sens."},{"key":"ref_11","doi-asserted-by":"crossref","first-page":"139","DOI":"10.1016\/j.isprsjprs.2017.05.002","article-title":"Simultaneous extraction of roads and buildings in remote sensing imagery with convolutional neural networks","volume":"130","author":"Alshehhi","year":"2017","journal-title":"ISPRS J. Photogramm. Remote Sens."},{"key":"ref_12","doi-asserted-by":"crossref","unstructured":"Long, J., Shelhamer, E., and Darrell, T. (2015, January 7\u201312). Fully convolutional networks for semantic segmentation. Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition, Boston, MA, USA.","DOI":"10.1109\/CVPR.2015.7298965"},{"key":"ref_13","doi-asserted-by":"crossref","unstructured":"Zhong, Z., Li, J., Cui, W., and Jiang, H. (2016, January 10\u201315). Fully convolutional networks for building and road extraction: Preliminary results. Proceedings of the IEEE International Geoscience and Remote Sensing Symposium (IGARSS), Beijing, China.","DOI":"10.1109\/IGARSS.2016.7729406"},{"key":"ref_14","doi-asserted-by":"crossref","first-page":"709","DOI":"10.1109\/LGRS.2017.2672734","article-title":"Road structure refined CNN for road extraction in aerial image","volume":"14","author":"Wei","year":"2017","journal-title":"IEEE Geosci. Remote Sens. Lett."},{"key":"ref_15","doi-asserted-by":"crossref","unstructured":"Ronneberger, O., Fischer, P., and Brox, T. (2015, January 5\u20139). U-Net: Convolutional networks for biomedical image segmentation. Proceedings of the Medical Image Computing and Computer-Assisted Intervention, Munich, Germany.","DOI":"10.1007\/978-3-319-24574-4_28"},{"key":"ref_16","doi-asserted-by":"crossref","first-page":"3322","DOI":"10.1109\/TGRS.2017.2669341","article-title":"Automatic road detection and centerline extraction via cascaded End-to-End convolutional neural network","volume":"55","author":"Cheng","year":"2017","journal-title":"IEEE Trans. Geosci. Remote Sens."},{"key":"ref_17","doi-asserted-by":"crossref","unstructured":"Panboonyuen, T., Jitkajornwanich, K., Lawawirojwong, S., Srestasathiern, P., and Vateekul, P. (2017). Road segmentation of remotely-Sensed images using deep convolutional neural networks with landscape metrics and conditional random fields. Remote Sens., 9.","DOI":"10.20944\/preprints201706.0012.v1"},{"key":"ref_18","unstructured":"Badrinarayanan, V., Kendall, A., and Cipolla, R. (2015). Segnet: A Deep convolutional encoder-decoder architecture for image segmentation. arXiv."},{"key":"ref_19","unstructured":"Clevert, D.A., Unterthiner, T., and Hochreiter, S. (2015). Fast and accurate deep network learning by exponential linear units (ELUs). arXiv."},{"key":"ref_20","doi-asserted-by":"crossref","first-page":"749","DOI":"10.1109\/LGRS.2018.2802944","article-title":"Road extraction by deep residual U-Net","volume":"15","author":"Zhang","year":"2018","journal-title":"IEEE Geosci. Remote Sens. Lett."},{"key":"ref_21","doi-asserted-by":"crossref","unstructured":"He, K., Zhang, X., Ren, S., and Sun, J. (2016, January 27\u201330). Deep residual learning for image recognition. Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition, Las Vegas, NV, USA.","DOI":"10.1109\/CVPR.2016.90"},{"key":"ref_22","doi-asserted-by":"crossref","unstructured":"Mosinska, A., M\u00e1rquez-Neila, P., Kozinski, M., and Fua, P. (2018, January 8\u201314). Beyond the pixel-wise loss for topology-aware delineation. Proceedings of the 2018 IEEE\/CVF Conference on Computer Vision and Pattern Recognition, Munich, Germany.","DOI":"10.1109\/CVPR.2018.00331"},{"key":"ref_23","unstructured":"Simonyan, K., and Zisserman, A. (2015, January 7\u20139). Very deep convolutional networks for large-scale image recognition. Proceedings of the International Conference on Learning Representations, San Diego, CA, USA."},{"key":"ref_24","unstructured":"Chen, L.C., Papandreou, G., Schroff, F., and Adam, H. (2017). Rethinking atrous convolution for semantic image segmentation. arXiv."},{"key":"ref_25","doi-asserted-by":"crossref","first-page":"600","DOI":"10.1109\/TIP.2003.819861","article-title":"Image quality assessment: from error visibility to structural similarity","volume":"13","author":"Zhou","year":"2004","journal-title":"IEEE Trans. Image Process."},{"key":"ref_26","first-page":"1","article-title":"Applicability of Existing Objective Metrics of Perceptual Quality for Adaptive Video Streaming","volume":"13","author":"Krasula","year":"2016","journal-title":"Electron. Imaging"},{"key":"ref_27","doi-asserted-by":"crossref","unstructured":"Shrestha, S., and Vanneschi, L. (2018). Improved Fully convolutional network with conditional random fields for building extraction. Remote Sens., 10.","DOI":"10.3390\/rs10071135"},{"key":"ref_28","unstructured":"Ioffe, S., and Szegedy, C. (2015, January 6\u201311). Batch normalization: Accelerating deep network training by reducing internal covariate shift. Proceedings of the Iernational Conference on Machine Learning, Lille, France."},{"key":"ref_29","doi-asserted-by":"crossref","unstructured":"Chen, L.C., Zhu, Y., Papandreou, G., Schroff, F., and Adam, H. (2018). Encoder-Decoder with atrous separable convolution for semantic image segmentation. arXiv.","DOI":"10.1007\/978-3-030-01234-2_49"},{"key":"ref_30","unstructured":"Iglovikov, V., Mushinskiy, S., and Osin, V. (2017). Satellite Imagery Feature Detection using Deep Convolutional Neural Network: A Kaggle Competition. arXiv."},{"key":"ref_31","first-page":"330","article-title":"A road extraction method for remote sensing image based on Encoder-Decoder network","volume":"48","author":"He","year":"2019","journal-title":"Acta Geod. Cartogr. Sin."},{"key":"ref_32","unstructured":"Goodfellow, I., Bengio, Y., and Courville, A. (2016). Deep Learning, MIT Press."},{"key":"ref_33","unstructured":"Wiedemann, C., Heipke, C., Mayer, H., and Jamet, O. (1998). Empirical Evaluation of Automatically Extracted Road Axes. Empirical Evaluation Techniques in Computer Vision, IEEE Computer Society Press."},{"key":"ref_34","unstructured":"Kingma, D., and Ba, J. (2014). Adam: A method for stochastic optimization. arXiv."},{"key":"ref_35","unstructured":"Ruder, S. (2016). An overview of gradient descent optimization. arXiv."}],"container-title":["Remote Sensing"],"original-title":[],"language":"en","link":[{"URL":"https:\/\/www.mdpi.com\/2072-4292\/11\/9\/1015\/pdf","content-type":"unspecified","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2025,10,11]],"date-time":"2025-10-11T12:47:56Z","timestamp":1760186876000},"score":1,"resource":{"primary":{"URL":"https:\/\/www.mdpi.com\/2072-4292\/11\/9\/1015"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2019,4,29]]},"references-count":35,"journal-issue":{"issue":"9","published-online":{"date-parts":[[2019,5]]}},"alternative-id":["rs11091015"],"URL":"https:\/\/doi.org\/10.3390\/rs11091015","relation":{},"ISSN":["2072-4292"],"issn-type":[{"value":"2072-4292","type":"electronic"}],"subject":[],"published":{"date-parts":[[2019,4,29]]}}}