{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2026,7,22]],"date-time":"2026-07-22T16:42:40Z","timestamp":1784738560823,"version":"3.55.0"},"reference-count":40,"publisher":"MDPI AG","issue":"3","license":[{"start":{"date-parts":[[2025,3,2]],"date-time":"2025-03-02T00:00:00Z","timestamp":1740873600000},"content-version":"vor","delay-in-days":0,"URL":"https:\/\/creativecommons.org\/licenses\/by\/4.0\/"}],"funder":[{"name":"Postgraduate Research &amp; Practice Innovation Program of Jiangsu Province","award":["SJCX24_1327"],"award-info":[{"award-number":["SJCX24_1327"]}]}],"content-domain":{"domain":[],"crossmark-restriction":false},"short-container-title":["Algorithms"],"abstract":"<jats:p>Abandoned objects on highways seriously threaten traffic safety, and their prompt identification and removal are crucial. Existing methods struggle to balance computational cost and detection accuracy due to the significant scale differences of abandoned objects on highways. To address these problems, we propose a Lightweight and Efficient Detection Transformer for highway abandoned objects (LE-DETR). This study first designs a real-time feature extraction module that effectively captures essential information and accelerates information flow. Building on this module, we construct a lightweight backbone network for feature extraction, enhancing parameter utilization. A Triple Fusion (TFusion) module is proposed, integrating high-level semantic information with low-level spatial information to increase detailed information. A Cross-Layer Multi-Scale Interaction (CMI) module is designed, utilizing large-kernel depth-wise convolutions of various sizes to extract features from different receptive fields, enhancing the multi-scale representation of abandoned objects. The LE-DETR model is trained and evaluated using a constructed Highway Abandoned Object Dataset (HAOD). The experimental results indicate that compared to the suboptimal RT-DETR-R18, LE-DETR improves accuracy by 6.5%, reduces the number of parameters by 27.1%, and decreases floating-point operations (FLOPs) by 21.1%. These improvements demonstrate the great potential of LE-DETR for detecting abandoned objects on highways.<\/jats:p>","DOI":"10.3390\/a18030133","type":"journal-article","created":{"date-parts":[[2025,3,3]],"date-time":"2025-03-03T03:22:44Z","timestamp":1740972164000},"page":"133","update-policy":"https:\/\/doi.org\/10.3390\/mdpi_crossmark_policy","source":"Crossref","is-referenced-by-count":5,"title":["A Lightweight and Efficient Detection Transformer for Highway Abandoned Objects"],"prefix":"10.3390","volume":"18","author":[{"given":"Biao","family":"Zhang","sequence":"first","affiliation":[{"name":"School of Network and Communication Engineering, Jinling Institute of Technology, Nanjing 211169, China"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Chishe","family":"Wang","sequence":"additional","affiliation":[{"name":"School of Network and Communication Engineering, Jinling Institute of Technology, Nanjing 211169, China"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"ORCID":"https:\/\/orcid.org\/0000-0003-3679-8678","authenticated-orcid":false,"given":"Jie","family":"Wang","sequence":"additional","affiliation":[{"name":"School of Network and Communication Engineering, Jinling Institute of Technology, Nanjing 211169, China"}],"role":[{"vocabulary":"crossref","role":"author"}]}],"member":"1968","published-online":{"date-parts":[[2025,3,2]]},"reference":[{"key":"ref_1","unstructured":"Fu, H., Xiang, M., Ma, H., Ming, A., and Liu, L. (2011, January 26\u201328). Abandoned object detection in highway scene. Proceedings of the 2011 6th International Conference on Pervasive Computing and Applications, Port Elizabeth, South Africa."},{"key":"ref_2","doi-asserted-by":"crossref","unstructured":"Zivkovic, Z. (2004, January 26\u201326). Improved adaptive Gaussian mixture model for background subtraction. Proceedings of the 17th International Conference on Pattern Recognition, 2004, ICPR 2004, Cambridge, UK.","DOI":"10.1109\/ICPR.2004.1333992"},{"key":"ref_3","doi-asserted-by":"crossref","first-page":"761","DOI":"10.1007\/s11042-014-2324-4","article-title":"Extraction of stable foreground image regions for unattended luggage detection","volume":"75","author":"Szwoch","year":"2016","journal-title":"Multimed. Tools Appl."},{"key":"ref_4","doi-asserted-by":"crossref","unstructured":"Zeng, Y., Lan, J., Ran, B., Gao, J., and Zou, J. (2015). A Novel Abandoned Object Detection System Based on Three-Dimensional Image Information. Sensors, 15.","DOI":"10.3390\/s150306885"},{"key":"ref_5","doi-asserted-by":"crossref","first-page":"181","DOI":"10.1016\/j.jvcir.2016.05.024","article-title":"Robust techniques for abandoned and removed object detection based on Markov random field","volume":"39","author":"Lin","year":"2016","journal-title":"J. Vis. Commun. Image Represent."},{"key":"ref_6","doi-asserted-by":"crossref","first-page":"80010","DOI":"10.1109\/ACCESS.2020.2990618","article-title":"Detection of Abandoned and Stolen Objects Based on Dual Background Model and Mask R-CNN","volume":"8","author":"Park","year":"2020","journal-title":"IEEE Access"},{"key":"ref_7","doi-asserted-by":"crossref","unstructured":"He, K., Gkioxari, G., Doll\u00e1r, P., and Girshick, R. (2018). Mask R-CNN. arXiv.","DOI":"10.1109\/ICCV.2017.322"},{"key":"ref_8","doi-asserted-by":"crossref","first-page":"54","DOI":"10.1007\/s11760-024-03609-z","article-title":"Detection of abandoned objects based on Yolov9 and background differencing","volume":"19","author":"Song","year":"2024","journal-title":"Signal Image Video Process."},{"key":"ref_9","doi-asserted-by":"crossref","first-page":"773","DOI":"10.1016\/j.patrec.2005.11.005","article-title":"Efficient adaptive density estimation per image pixel for the task of background subtraction","volume":"27","author":"Zivkovic","year":"2006","journal-title":"Pattern Recognit. Lett."},{"key":"ref_10","doi-asserted-by":"crossref","unstructured":"Wang, C.-Y., Yeh, I.-H., and Liao, H.-Y.M. (2024). YOLOv9: Learning What You Want to Learn Using Programmable Gradient Information. arXiv.","DOI":"10.1007\/978-3-031-72751-1_1"},{"key":"ref_11","doi-asserted-by":"crossref","first-page":"1817","DOI":"10.1109\/TITS.2020.3015530","article-title":"Residual-Network-Leveraged Vehicle-Thrown-Waste Identification in Real-Time Traffic Surveillance Videos","volume":"22","author":"Qian","year":"2021","journal-title":"IEEE Trans. Intell. Transp. Syst."},{"key":"ref_12","doi-asserted-by":"crossref","first-page":"154","DOI":"10.1007\/s11263-013-0620-5","article-title":"Selective Search for Object Recognition","volume":"104","author":"Uijlings","year":"2013","journal-title":"Int. J. Comput. Vis."},{"key":"ref_13","doi-asserted-by":"crossref","unstructured":"van de Sande, K.E.A., Uijlings, J.R.R., Gevers, T., and Smeulders, A.W.M. (2011, January 6\u201313). Segmentation as selective search for object recognition. Proceedings of the 2011 International Conference on Computer Vision, Barcelona, Spain.","DOI":"10.1109\/ICCV.2011.6126456"},{"key":"ref_14","doi-asserted-by":"crossref","first-page":"237","DOI":"10.1080\/2150704X.2017.1415473","article-title":"Accurate non-maximum suppression for object detection in high-resolution remote sensing images","volume":"9","author":"Qiu","year":"2018","journal-title":"Remote Sens. Lett."},{"key":"ref_15","doi-asserted-by":"crossref","unstructured":"Neubeck, A., and Van Gool, L. (2006, January 20\u201324). Efficient Non-Maximum Suppression. Proceedings of the 18th International Conference on Pattern Recognition (ICPR\u201906), Hong Kong, China.","DOI":"10.1109\/ICPR.2006.479"},{"key":"ref_16","doi-asserted-by":"crossref","unstructured":"Liu, W., Li, J., Liu, Y., Ma, C., Xu, J., and Liu, C. (2023, January 23\u201325). Research on the lightweight detection network of abandoned objects in freeway based on video. Proceedings of the Sixth International Conference on Traffic Engineering and Transportation System (ICTETS 2022), Guangzhou, China.","DOI":"10.1117\/12.2668510"},{"key":"ref_17","doi-asserted-by":"crossref","unstructured":"Zhou, L., and Xu, J. (2024). Enhanced Abandoned Object Detection through Adaptive Dual-Background Modeling and SAO-YOLO Integration. Sensors, 24.","DOI":"10.3390\/s24206572"},{"key":"ref_18","doi-asserted-by":"crossref","unstructured":"Zhao, Y., Lv, W., Xu, S., Wei, J., Wang, G., Dang, Q., Liu, Y., and Chen, J. (2024, January 17\u201321). DETRs Beat YOLOs on Real-time Object Detection. Proceedings of the 2024 IEEE\/CVF Conference on Computer Vision and Pattern Recognition (CVPR), Seattle, WA, USA.","DOI":"10.1109\/CVPR52733.2024.01605"},{"key":"ref_19","unstructured":"Vaswani, A., Shazeer, N., Parmar, N., Uszkoreit, J., Jones, L., Gomez, A.N., Kaiser, L., and Polosukhin, I. (2023). Attention Is All You Need. arXiv."},{"key":"ref_20","doi-asserted-by":"crossref","unstructured":"Ding, X., Zhang, X., Ma, N., Han, J., Ding, G., and Sun, J. (2021, January 19\u201325). RepVGG: Making VGG-style ConvNets Great Again. Proceedings of the 2021 IEEE\/CVF Conference on Computer Vision and Pattern Recognition (CVPR), Nashville, TN, USA.","DOI":"10.1109\/CVPR46437.2021.01352"},{"key":"ref_21","unstructured":"Simonyan, K., and Zisserman, A. (2015). Very Deep Convolutional Networks for Large-Scale Image Recognition. arXiv."},{"key":"ref_22","doi-asserted-by":"crossref","unstructured":"Han, K., Wang, Y., Tian, Q., Guo, J., Xu, C., and Xu, C. (2020, January 13\u201319). GhostNet: More Features From Cheap Operations. Proceedings of the 2020 IEEE\/CVF Conference on Computer Vision and Pattern Recognition (CVPR), Seattle, WA, USA.","DOI":"10.1109\/CVPR42600.2020.00165"},{"key":"ref_23","doi-asserted-by":"crossref","first-page":"105057","DOI":"10.1016\/j.imavis.2024.105057","article-title":"ASF-YOLO: A novel YOLO model with attentional scale sequence fusion for cell instance segmentation","volume":"147","author":"Kang","year":"2024","journal-title":"Image Vis. Comput."},{"key":"ref_24","unstructured":"Chen, Y., Zhang, P., Li, Z., Li, Y., Zhang, X., Qi, L., Sun, J., and Jia, J. (2021). Dynamic Scale Training for Object Detection. arXiv."},{"key":"ref_25","first-page":"1","article-title":"Rotation Equivariant Feature Image Pyramid Network for Object Detection in Optical Remote Sensing Imagery","volume":"60","author":"Shamsolmoali","year":"2022","journal-title":"IEEE Trans. Geosci. Remote Sens."},{"key":"ref_26","doi-asserted-by":"crossref","unstructured":"Zhou, Y., Yang, X., Zhang, G., Wang, J., Liu, Y., Hou, L., Jiang, X., Liu, X., Yan, J., and Lyu, C. (2022, January 10\u201314). MMRotate: A Rotated Object Detection Benchmark using PyTorch. Proceedings of the 30th ACM International Conference on Multimedia, in MM \u201922, Lisbon, Portugal.","DOI":"10.1145\/3503161.3548541"},{"key":"ref_27","doi-asserted-by":"crossref","first-page":"5535","DOI":"10.1109\/TGRS.2019.2900302","article-title":"Hierarchical and Robust Convolutional Neural Network for Very High-Resolution Remote Sensing Object Detection","volume":"57","author":"Zhang","year":"2019","journal-title":"IEEE Trans. Geosci. Remote Sens."},{"key":"ref_28","doi-asserted-by":"crossref","first-page":"751","DOI":"10.1109\/LGRS.2018.2882551","article-title":"Squeeze and Excitation Rank Faster R-CNN for Ship Detection in SAR Images","volume":"16","author":"Lin","year":"2019","journal-title":"IEEE Geosci. Remote Sens. Lett."},{"key":"ref_29","doi-asserted-by":"crossref","first-page":"1758","DOI":"10.1109\/TCSVT.2019.2905881","article-title":"Small Object Detection in Unmanned Aerial Vehicle Images Using Feature Fusion and Scaling-Based Single Shot Detector With Spatial Context Analysis","volume":"30","author":"Liang","year":"2020","journal-title":"IEEE Trans. Circuits Syst. Video Technol."},{"key":"ref_30","doi-asserted-by":"crossref","first-page":"5602918","DOI":"10.1109\/TGRS.2022.3224815","article-title":"FSoD-Net: Full-Scale Object Detection From Optical Remote Sensing Imagery","volume":"60","author":"Wang","year":"2022","journal-title":"IEEE Trans. Geosci. Remote Sens."},{"key":"ref_31","unstructured":"Wang, A., Chen, H., Liu, L., Chen, K., Lin, Z., and Han, J. (2024). Yolov10: Real-time end-to-end object detection. arXiv."},{"key":"ref_32","doi-asserted-by":"crossref","unstructured":"Szegedy, C., Vanhoucke, V., Ioffe, S., Shlens, J., and Wojna, Z. (2016, January 27\u201330). Rethinking the Inception Architecture for Computer Vision. Proceedings of the 2016 IEEE Conference on Computer Vision and Pattern Recognition (CVPR), Las Vegas, NV, USA.","DOI":"10.1109\/CVPR.2016.308"},{"key":"ref_33","doi-asserted-by":"crossref","unstructured":"Yu, W., Zhou, P., Yan, S., and Wang, X. (2024, January 17\u201321). InceptionNeXt: When Inception Meets ConvNeXt. Proceedings of the 2024 IEEE\/CVF Conference on Computer Vision and Pattern Recognition (CVPR), Seattle, WA, USA.","DOI":"10.1109\/CVPR52733.2024.00542"},{"key":"ref_34","doi-asserted-by":"crossref","unstructured":"Cai, X., Lai, Q., Wang, Y., Wang, W., Sun, Z., and Yao, Y. (2024, January 17\u201321). Poly Kernel Inception Network for Remote Sensing Detection. Proceedings of the 2024 IEEE\/CVF Conference on Computer Vision and Pattern Recognition (CVPR), Seattle, WA, USA.","DOI":"10.1109\/CVPR52733.2024.02617"},{"key":"ref_35","unstructured":"Xu, S., Wang, X., Lv, W., Chang, Q., Cui, C., Deng, K., Wang, G., Dang, Q., Wei, S., and Du, Y. (2022). PP-YOLOE: An evolved version of YOLO. arXiv."},{"key":"ref_36","doi-asserted-by":"crossref","unstructured":"Wang, C.-Y., Bochkovskiy, A., and Liao, H.-Y.M. (2023, January 18\u201322). YOLOv7: Trainable Bag-of-Freebies Sets New State-of-the-Art for Real-Time Object Detectors. Proceedings of the 2023 IEEE\/CVF Conference on Computer Vision and Pattern Recognition (CVPR), Vancouver, BC, Canada.","DOI":"10.1109\/CVPR52729.2023.00721"},{"key":"ref_37","doi-asserted-by":"crossref","first-page":"1137","DOI":"10.1109\/TPAMI.2016.2577031","article-title":"Faster R-CNN: Towards Real-Time Object Detection with Region Proposal Networks","volume":"39","author":"Ren","year":"2017","journal-title":"IEEE Trans. Pattern Anal. Mach. Intell."},{"key":"ref_38","doi-asserted-by":"crossref","unstructured":"Vedaldi, A., Bischof, H., Brox, T., and Frahm, J.-M. (2020). End-to-End Object Detection with Transformers. Computer Vision\u2014ECCV 2020, Springer International Publishing.","DOI":"10.1007\/978-3-030-58583-9"},{"key":"ref_39","unstructured":"Zhu, X., Su, W., Lu, L., Li, B., Wang, X., and Dai, J. (2021). Deformable DETR: Deformable Transformers for End-to-End Object Detection. arXiv."},{"key":"ref_40","unstructured":"Zhang, H., Li, F., Liu, S., Zhang, L., Su, H., Zhu, J., Ni, L.M., and Shum, H.-Y. (2022). DINO: DETR with Improved DeNoising Anchor Boxes for End-to-End Object Detection. arXiv."}],"container-title":["Algorithms"],"original-title":[],"language":"en","link":[{"URL":"https:\/\/www.mdpi.com\/1999-4893\/18\/3\/133\/pdf","content-type":"unspecified","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2025,10,9]],"date-time":"2025-10-09T16:45:51Z","timestamp":1760028351000},"score":1,"resource":{"primary":{"URL":"https:\/\/www.mdpi.com\/1999-4893\/18\/3\/133"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2025,3,2]]},"references-count":40,"journal-issue":{"issue":"3","published-online":{"date-parts":[[2025,3]]}},"alternative-id":["a18030133"],"URL":"https:\/\/doi.org\/10.3390\/a18030133","relation":{},"ISSN":["1999-4893"],"issn-type":[{"value":"1999-4893","type":"electronic"}],"subject":[],"published":{"date-parts":[[2025,3,2]]}}}