{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2026,6,27]],"date-time":"2026-06-27T16:03:33Z","timestamp":1782576213636,"version":"3.54.5"},"reference-count":54,"publisher":"MDPI AG","issue":"6","license":[{"start":{"date-parts":[[2024,3,11]],"date-time":"2024-03-11T00:00:00Z","timestamp":1710115200000},"content-version":"vor","delay-in-days":0,"URL":"https:\/\/creativecommons.org\/licenses\/by\/4.0\/"}],"funder":[{"DOI":"10.13039\/501100012166","name":"National Key R&amp;D Program of China","doi-asserted-by":"publisher","award":["2022YFC3004701"],"award-info":[{"award-number":["2022YFC3004701"]}],"id":[{"id":"10.13039\/501100012166","id-type":"DOI","asserted-by":"publisher"}]}],"content-domain":{"domain":[],"crossmark-restriction":false},"short-container-title":["Remote Sensing"],"abstract":"<jats:p>Pavement crack detection is of significant importance in ensuring road safety and smooth traffic flow. However, pavement cracks come in various shapes and forms which exhibit spatial continuity, and algorithms need to adapt to different types of cracks while preserving their continuity. To address these challenges, an innovative crack detection framework, CrackDiff, based on the generative diffusion model, is proposed. It leverages the learning capabilities of the generative diffusion model for the data distribution and latent spatial relationships of cracks across different sample timesteps and generates more accurate and continuous crack segmentation results. CrackDiff uses crack images as guidance for the diffusion model and employs a multi-task UNet architecture to predict mask and noise simultaneously at each sampling step, enhancing the robustness of generations. Compared to other models, CrackDiff generates more accurate and stable results. Through experiments on the Crack500 and DeepCrack pavement datasets, CrackDiff achieves the best performance (F1 = 0.818 and mIoU = 0.841 on Crack500, and F1 = 0.841 and mIoU = 0.862 on DeepCrack).<\/jats:p>","DOI":"10.3390\/rs16060986","type":"journal-article","created":{"date-parts":[[2024,3,12]],"date-time":"2024-03-12T03:55:34Z","timestamp":1710215734000},"page":"986","update-policy":"https:\/\/doi.org\/10.3390\/mdpi_crossmark_policy","source":"Crossref","is-referenced-by-count":37,"title":["The Crack Diffusion Model: An Innovative Diffusion-Based Method for Pavement Crack Detection"],"prefix":"10.3390","volume":"16","author":[{"ORCID":"https:\/\/orcid.org\/0000-0001-7978-8045","authenticated-orcid":false,"given":"Haoyuan","family":"Zhang","sequence":"first","affiliation":[{"name":"Institude of Remote Sensing and Geographic Information System, Peking University, Beijing 100871, China"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"ORCID":"https:\/\/orcid.org\/0009-0001-9280-975X","authenticated-orcid":false,"given":"Ning","family":"Chen","sequence":"additional","affiliation":[{"name":"Institude of Remote Sensing and Geographic Information System, Peking University, Beijing 100871, China"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"ORCID":"https:\/\/orcid.org\/0000-0002-4763-1817","authenticated-orcid":false,"given":"Mei","family":"Li","sequence":"additional","affiliation":[{"name":"Institude of Remote Sensing and Geographic Information System, Peking University, Beijing 100871, China"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Shanjun","family":"Mao","sequence":"additional","affiliation":[{"name":"Institude of Remote Sensing and Geographic Information System, Peking University, Beijing 100871, China"}],"role":[{"vocabulary":"crossref","role":"author"}]}],"member":"1968","published-online":{"date-parts":[[2024,3,11]]},"reference":[{"key":"ref_1","doi-asserted-by":"crossref","first-page":"1525","DOI":"10.1109\/TITS.2019.2910595","article-title":"Feature pyramid and hierarchical boosting network for pavement crack detection","volume":"21","author":"Yang","year":"2019","journal-title":"IEEE Trans. Intell. Transp. Syst."},{"key":"ref_2","doi-asserted-by":"crossref","first-page":"182060","DOI":"10.1109\/ACCESS.2019.2958264","article-title":"Infrared Thermal Imaging-Based Crack Detection Using Deep Learning","volume":"7","author":"Yang","year":"2019","journal-title":"IEEE Access"},{"key":"ref_3","doi-asserted-by":"crossref","unstructured":"Li, Q., Zhang, D., Zou, Q., and Lin, H. (September, January 28). 3D laser imaging and sparse points grouping for pavement crack detection. Proceedings of the 2017 25th European Signal Processing Conference (EUSIPCO), Kos, Greece.","DOI":"10.23919\/EUSIPCO.2017.8081567"},{"key":"ref_4","doi-asserted-by":"crossref","first-page":"126162","DOI":"10.1016\/j.conbuildmat.2021.126162","article-title":"A critical review and comparative study on image segmentation-based techniques for pavement crack detection","volume":"321","author":"Kheradmandi","year":"2022","journal-title":"Constr. Build. Mater."},{"key":"ref_5","doi-asserted-by":"crossref","unstructured":"Zhao, H., Qin, G., and Wang, X. (2010, January 16\u201318). Improvement of canny algorithm based on pavement edge detection. Proceedings of the 2010 3rd International Congress on Image and Signal Processing, Yantai, China.","DOI":"10.1109\/CISP.2010.5646923"},{"key":"ref_6","doi-asserted-by":"crossref","first-page":"7163580","DOI":"10.1155\/2018\/7163580","article-title":"Metaheuristic optimized edge detection for recognition of concrete wall cracks: A comparative study on the performances of roberts, prewitt, canny, and sobel algorithms","volume":"2018","author":"Hoang","year":"2018","journal-title":"Adv. Civ. Eng."},{"key":"ref_7","first-page":"103527","article-title":"Robust surface crack detection with structure line guidance","volume":"124","author":"Zhang","year":"2023","journal-title":"Int. J. Appl. Earth Obs. Geoinf."},{"key":"ref_8","doi-asserted-by":"crossref","unstructured":"Lin, J., and Liu, Y. (2010, January 10\u201312). Potholes detection based on SVM in the pavement distress image. Proceedings of the International Symposium DCABES, Hong Kong, China.","DOI":"10.1109\/DCABES.2010.115"},{"key":"ref_9","doi-asserted-by":"crossref","first-page":"162","DOI":"10.1111\/j.1467-8667.2012.00790.x","article-title":"Texture analysis based damage detection of ageing infrastructural elements","volume":"28","author":"Schoefs","year":"2013","journal-title":"Comput.-Aided Civ. Infrastruct. Eng."},{"key":"ref_10","doi-asserted-by":"crossref","first-page":"3434","DOI":"10.1109\/TITS.2016.2552248","article-title":"Automatic road crack detection using random structured forests","volume":"17","author":"Shi","year":"2016","journal-title":"IEEE Trans. Intell. Transp. Syst."},{"key":"ref_11","doi-asserted-by":"crossref","unstructured":"David Jenkins, M., Carr, T.A., Iglesias, M.I., Buggy, T., and Morison, G. (2018, January 3\u20137). A Deep Convolutional Neural Network for Semantic Pixel-Wise Segmentation of Road and Pavement Surface Cracks. Proceedings of the 2018 26th European Signal Processing Conference (EUSIPCO), Roma, Italy.","DOI":"10.23919\/EUSIPCO.2018.8553280"},{"key":"ref_12","doi-asserted-by":"crossref","first-page":"139","DOI":"10.1016\/j.neucom.2019.01.036","article-title":"DeepCrack: A deep hierarchical feature learning architecture for crack segmentation","volume":"338","author":"Liu","year":"2019","journal-title":"Neurocomputing"},{"key":"ref_13","doi-asserted-by":"crossref","first-page":"04019040","DOI":"10.1061\/(ASCE)CP.1943-5487.0000854","article-title":"Robust pixel-level crack detection using deep fully convolutional neural networks","volume":"33","author":"Alipour","year":"2019","journal-title":"J. Comput. Civ. Eng."},{"key":"ref_14","first-page":"100144","article-title":"Pavement crack detection and recognition using the architecture of segNet","volume":"18","author":"Chen","year":"2020","journal-title":"J. Ind. Inf. Integr."},{"key":"ref_15","doi-asserted-by":"crossref","first-page":"18392","DOI":"10.1109\/TITS.2022.3158670","article-title":"DMA-Net: DeepLab with multi-scale attention for pavement crack segmentation","volume":"23","author":"Sun","year":"2022","journal-title":"IEEE Trans. Intell. Transp. Syst."},{"key":"ref_16","doi-asserted-by":"crossref","first-page":"103176","DOI":"10.1016\/j.autcon.2020.103176","article-title":"An integrated approach to automatic pixel-level crack detection and quantification of asphalt pavement","volume":"114","author":"Ji","year":"2020","journal-title":"Autom. Constr."},{"key":"ref_17","first-page":"103039","article-title":"MFPA-Net: An efficient deep learning network for automatic ground fissures extraction in UAV images of the coal mining area","volume":"114","author":"Jiang","year":"2022","journal-title":"Int. J. Appl. Earth Obs. Geoinf."},{"key":"ref_18","first-page":"103172","article-title":"Pavement crack detection with hybrid-window attentive vision transformers","volume":"116","author":"Xiao","year":"2023","journal-title":"Int. J. Appl. Earth Obs. Geoinf."},{"key":"ref_19","doi-asserted-by":"crossref","first-page":"103989","DOI":"10.1016\/j.autcon.2021.103989","article-title":"Structural crack detection using deep convolutional neural networks","volume":"133","author":"Ali","year":"2022","journal-title":"Autom. Constr."},{"key":"ref_20","unstructured":"Ho, J., Jain, A., and Abbeel, P. (2020, January 6\u201312). Denoising Diffusion Probabilistic Models. Proceedings of the 34th International Conference on Neural Information Processing Systems, Red Hook, NY, USA."},{"key":"ref_21","doi-asserted-by":"crossref","first-page":"53","DOI":"10.1109\/MSP.2017.2765202","article-title":"Generative adversarial networks: An overview","volume":"35","author":"Creswell","year":"2018","journal-title":"IEEE Signal Process. Mag."},{"key":"ref_22","unstructured":"Kingma, D.P., and Welling, M. (2013). Auto-encoding variational bayes. arXiv."},{"key":"ref_23","doi-asserted-by":"crossref","unstructured":"Han, L., Zhao, Y., Lv, H., Zhang, Y., Liu, H., Bi, G., and Han, Q. (2023). Enhancing Remote Sensing Image Super-Resolution with Efficient Hybrid Conditional Diffusion Model. Remote Sens., 15.","DOI":"10.3390\/rs15133452"},{"key":"ref_24","doi-asserted-by":"crossref","unstructured":"Ronneberger, O., Fischer, P., and Brox, T. (2015, January 5\u20139). U-Net: Convolutional Networks for Biomedical Image Segmentation. Proceedings of the Medical Image Computing and Computer-Assisted Intervention\u2014MICCAI 2015, Munich, Germany.","DOI":"10.1007\/978-3-319-24574-4_28"},{"key":"ref_25","doi-asserted-by":"crossref","first-page":"640","DOI":"10.1109\/TPAMI.2016.2572683","article-title":"Fully Convolutional Networks for Semantic Segmentation","volume":"39","author":"Shelhamer","year":"2017","journal-title":"IEEE Trans. Pattern Anal. Mach. Intell."},{"key":"ref_26","doi-asserted-by":"crossref","unstructured":"Chen, L.C., Zhu, Y., Papandreou, G., Schroff, F., and Adam, H. (2018, January 8\u201314). Encoder-decoder with atrous separable convolution for semantic image segmentation. Proceedings of the Computer Vision\u2014ECCV 2018: 15th European Conference, Berlin\/Heidelberg, Germany.","DOI":"10.1007\/978-3-030-01234-2_49"},{"key":"ref_27","doi-asserted-by":"crossref","first-page":"2481","DOI":"10.1109\/TPAMI.2016.2644615","article-title":"Segnet: A deep convolutional encoder-decoder architecture for image segmentation","volume":"39","author":"Badrinarayanan","year":"2017","journal-title":"IEEE Trans. Pattern Anal. Mach. Intell."},{"key":"ref_28","doi-asserted-by":"crossref","unstructured":"Zhao, H., Shi, J., Qi, X., Wang, X., and Jia, J. (2017, January 21\u201326). Pyramid scene parsing network. Proceedings of the 2017 IEEE Conference on Computer Vision and Pattern Recognition (CVPR), Los Alamitos, CA, USA.","DOI":"10.1109\/CVPR.2017.660"},{"key":"ref_29","first-page":"103335","article-title":"YOLOv5s-M: A deep learning network model for road pavement damage detection from urban street-view imagery","volume":"120","author":"Ren","year":"2023","journal-title":"Int. J. Appl. Earth Obs. Geoinf."},{"key":"ref_30","doi-asserted-by":"crossref","first-page":"6412562","DOI":"10.1155\/2020\/6412562","article-title":"Automated pavement crack damage detection using deep multiscale convolutional features","volume":"2020","author":"Song","year":"2020","journal-title":"J. Adv. Transp."},{"key":"ref_31","doi-asserted-by":"crossref","first-page":"3349","DOI":"10.1109\/TPAMI.2020.2983686","article-title":"Deep high-resolution representation learning for visual recognition","volume":"43","author":"Wang","year":"2020","journal-title":"IEEE Trans. Pattern Anal. Mach. Intell."},{"key":"ref_32","doi-asserted-by":"crossref","first-page":"103357","DOI":"10.1016\/j.autcon.2020.103357","article-title":"A spatial-channel hierarchical deep learning network for pixel-level automated crack detection","volume":"119","author":"Pan","year":"2020","journal-title":"Autom. Constr."},{"key":"ref_33","doi-asserted-by":"crossref","first-page":"1859","DOI":"10.1177\/1369433220986638","article-title":"Intelligent crack detection based on attention mechanism in convolution neural network","volume":"24","author":"Cui","year":"2021","journal-title":"Adv. Struct. Eng"},{"key":"ref_34","doi-asserted-by":"crossref","first-page":"110216","DOI":"10.1016\/j.knosys.2022.110216","article-title":"Concrete crack detection using lightweight attention feature fusion single shot multibox detector","volume":"261","author":"Zhu","year":"2023","journal-title":"Knowl.-Based Syst."},{"key":"ref_35","first-page":"12077","article-title":"SegFormer: Simple and efficient design for semantic segmentation with transformers","volume":"34","author":"Xie","year":"2021","journal-title":"Proc. Adv. Neural Inf. Process. Syst."},{"key":"ref_36","doi-asserted-by":"crossref","unstructured":"Liu, H., Miao, X., Mertz, C., Xu, C., and Kong, H. (2021, January 11\u201317). Crackformer: Transformer network for fine-grained crack detection. Proceedings of the 2021 IEEE\/CVF International Conference on Computer Vision (ICCV), Montreal, QC, Canada.","DOI":"10.1109\/ICCV48922.2021.00376"},{"key":"ref_37","first-page":"102825","article-title":"Pavement crack detection from CCD images with a locally enhanced transformer network","volume":"110","author":"Xu","year":"2022","journal-title":"Int. J. Appl. Earth Obs. Geoinf."},{"key":"ref_38","doi-asserted-by":"crossref","unstructured":"Tao, H., Liu, B., Cui, J., and Zhang, H. (2023, January 9\u201312). A Convolutional-Transformer Network for Crack Segmentation with Boundary Awareness. Proceedings of the 2023 IEEE International Conference on Image Processing (ICIP), Kuala Lumpur, Malaysia.","DOI":"10.1109\/ICIP49359.2023.10222276"},{"key":"ref_39","doi-asserted-by":"crossref","first-page":"1306","DOI":"10.1109\/TITS.2020.2990703","article-title":"CrackGAN: Pavement Crack Detection Using Partially Accurate Ground Truths Based on Generative Adversarial Learning","volume":"22","author":"Zhang","year":"2021","journal-title":"IEEE Trans. Intell. Transp. Syst."},{"key":"ref_40","doi-asserted-by":"crossref","unstructured":"Liu, Y., Gao, W., Zhao, T., Wang, Z., and Wang, Z. (2023). A Rapid Bridge Crack Detection Method Based on Deep Learning. Appl. Sci., 13.","DOI":"10.3390\/app13179878"},{"key":"ref_41","doi-asserted-by":"crossref","unstructured":"Kyslytsyna, A., Xia, K., Kislitsyn, A., Abd El Kader, I., and Wu, Y. (2021). Road Surface Crack Detection Method Based on Conditional Generative Adversarial Networks. Sensors, 21.","DOI":"10.3390\/s21217405"},{"key":"ref_42","first-page":"1415","article-title":"Maximum likelihood training of score-based diffusion models","volume":"34","author":"Song","year":"2021","journal-title":"Adv. Neural Inf. Process. Syst."},{"key":"ref_43","unstructured":"Song, Y., Sohl-Dickstein, J., Kingma, D.P., Kumar, A., Ermon, S., and Poole, B. (2020). Score-based generative modeling through stochastic differential equations. arXiv."},{"key":"ref_44","unstructured":"Song, J., Meng, C., and Ermon, S. (2020). Denoising diffusion implicit models. arXiv."},{"key":"ref_45","unstructured":"Nichol, A.Q., and Dhariwal, P. (2021, January 18\u201324). Improved denoising diffusion probabilistic models. Proceedings of the 38th International Conference on Machine Learning, Virtual."},{"key":"ref_46","first-page":"12533","article-title":"D2C: Diffusion-decoding models for few-shot conditional generation","volume":"34","author":"Ranzato","year":"2021","journal-title":"Adv. Neural Inf. Process. Syst."},{"key":"ref_47","doi-asserted-by":"crossref","unstructured":"Peebles, W., and Xie, S. (2023, January 2\u20133). Scalable diffusion models with transformers. Proceedings of the IEEE\/CVF International Conference on Computer Vision (ICCV), Paris, France.","DOI":"10.1109\/ICCV51070.2023.00387"},{"key":"ref_48","doi-asserted-by":"crossref","first-page":"47","DOI":"10.1016\/j.neucom.2022.01.029","article-title":"Srdiff: Single image super-resolution with diffusion probabilistic models","volume":"479","author":"Li","year":"2022","journal-title":"Neurocomputing"},{"key":"ref_49","unstructured":"Amit, T., Shaharbany, T., Nachmani, E., and Wolf, L. (2021). Segdiff: Image segmentation with diffusion probabilistic models. arXiv."},{"key":"ref_50","doi-asserted-by":"crossref","unstructured":"Saharia, C., Chan, W., Chang, H., Lee, C., Ho, J., Salimans, T., Fleet, D., and Norouzi, M. (2022, January 7\u201311). Palette: Image-to-image diffusion models. Proceedings of the ACM SIGGRAPH 2022 Conference Proceedings, Vancouver, BC, Canada.","DOI":"10.1145\/3528233.3530757"},{"key":"ref_51","doi-asserted-by":"crossref","first-page":"5611317","DOI":"10.1109\/TGRS.2023.3279864","article-title":"PanDiff: A Novel Pansharpening Method Based on Denoising Diffusion Probabilistic Model","volume":"61","author":"Meng","year":"2023","journal-title":"IEEE Trans. Geosci. Remote Sens."},{"key":"ref_52","doi-asserted-by":"crossref","unstructured":"Zhang, Y., Tian, Y., Kong, Y., Zhong, B., and Fu, Y. (2018, January 18\u201322). Residual Dense Network for Image Super-Resolution. Proceedings of the 2018 IEEE\/CVF Conference on Computer Vision and Pattern Recognition, Salt Lake City, UT, USA.","DOI":"10.1109\/CVPR.2018.00262"},{"key":"ref_53","doi-asserted-by":"crossref","unstructured":"Wang, X., Girshick, R., Gupta, A., and He, K. (2018, January 18\u201322). Non-local neural networks. Proceedings of the 2018 IEEE\/CVF Conference on Computer Vision and Pattern Recognition, Salt Lake City, UT, USA.","DOI":"10.1109\/CVPR.2018.00813"},{"key":"ref_54","doi-asserted-by":"crossref","first-page":"227","DOI":"10.1016\/j.patrec.2011.11.004","article-title":"CrackTree: Automatic crack detection from pavement images","volume":"33","author":"Zou","year":"2012","journal-title":"Pattern Recognit. Lett."}],"container-title":["Remote Sensing"],"original-title":[],"language":"en","link":[{"URL":"https:\/\/www.mdpi.com\/2072-4292\/16\/6\/986\/pdf","content-type":"unspecified","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2025,10,10]],"date-time":"2025-10-10T14:12:06Z","timestamp":1760105526000},"score":1,"resource":{"primary":{"URL":"https:\/\/www.mdpi.com\/2072-4292\/16\/6\/986"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2024,3,11]]},"references-count":54,"journal-issue":{"issue":"6","published-online":{"date-parts":[[2024,3]]}},"alternative-id":["rs16060986"],"URL":"https:\/\/doi.org\/10.3390\/rs16060986","relation":{},"ISSN":["2072-4292"],"issn-type":[{"value":"2072-4292","type":"electronic"}],"subject":[],"published":{"date-parts":[[2024,3,11]]}}}