{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2026,6,5]],"date-time":"2026-06-05T00:35:54Z","timestamp":1780619754089,"version":"3.54.1"},"reference-count":38,"publisher":"MDPI AG","issue":"13","license":[{"start":{"date-parts":[[2023,6,27]],"date-time":"2023-06-27T00:00:00Z","timestamp":1687824000000},"content-version":"vor","delay-in-days":0,"URL":"https:\/\/creativecommons.org\/licenses\/by\/4.0\/"}],"funder":[{"name":"National Key R&amp;D Program of China","award":["2021YFB2501800"],"award-info":[{"award-number":["2021YFB2501800"]}]},{"name":"National Key R&amp;D Program of China","award":["52102461"],"award-info":[{"award-number":["52102461"]}]},{"name":"National Key R&amp;D Program of China","award":["32115013"],"award-info":[{"award-number":["32115013"]}]},{"name":"National Natural Science Foundation of China","award":["2021YFB2501800"],"award-info":[{"award-number":["2021YFB2501800"]}]},{"name":"National Natural Science Foundation of China","award":["52102461"],"award-info":[{"award-number":["52102461"]}]},{"name":"National Natural Science Foundation of China","award":["32115013"],"award-info":[{"award-number":["32115013"]}]},{"name":"State Key Laboratory of Advanced Design and Manufacturing for Vehicle Body","award":["2021YFB2501800"],"award-info":[{"award-number":["2021YFB2501800"]}]},{"name":"State Key Laboratory of Advanced Design and Manufacturing for Vehicle Body","award":["52102461"],"award-info":[{"award-number":["52102461"]}]},{"name":"State Key Laboratory of Advanced Design and Manufacturing for Vehicle Body","award":["32115013"],"award-info":[{"award-number":["32115013"]}]}],"content-domain":{"domain":[],"crossmark-restriction":false},"short-container-title":["Sensors"],"abstract":"<jats:p>To tackle the challenges posed by dense small objects and fuzzy boundaries on unstructured roads in the mining scenario, we proposed an end-to-end small object detection and drivable area segmentation framework for open-pit mining. We employed a convolutional network backbone as a feature extractor for both two tasks, as multi-task learning yielded promising results in autonomous driving perception. To address small object detection, we introduced a lightweight attention module that allowed our network to focus more on the spatial and channel dimensions of small objects without impeding inference time. We also used a convolutional block attention module in the drivable area segmentation subnetwork, which assigned more weight to road boundaries to improve feature mapping capabilities. Furthermore, to improve our network perception accuracy of both tasks, we used weighted summation when designing the loss function. We validated the effectiveness of our approach by testing it on pre-collected mining data which were called Minescape. Our detection results on the Minescape dataset showed 87.8% mAP index, which was 9.3% higher than state-of-the-art algorithms. Our segmentation results surpassed the comparison algorithm by 1 percent in MIoU index. Our experimental results demonstrated that our approach achieves competitive performance.<\/jats:p>","DOI":"10.3390\/s23135977","type":"journal-article","created":{"date-parts":[[2023,6,28]],"date-time":"2023-06-28T00:45:11Z","timestamp":1687913111000},"page":"5977","update-policy":"https:\/\/doi.org\/10.3390\/mdpi_crossmark_policy","source":"Crossref","is-referenced-by-count":11,"title":["MineSDS: A Unified Framework for Small Object Detection and Drivable Area Segmentation for Open-Pit Mining Scenario"],"prefix":"10.3390","volume":"23","author":[{"given":"Yong","family":"Liu","sequence":"first","affiliation":[{"name":"Zhuzhou CRRC Times Electric Co., Ltd., Zhuzhou 412001, China"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Cheng","family":"Li","sequence":"additional","affiliation":[{"name":"Zhuzhou CRRC Times Electric Co., Ltd., Zhuzhou 412001, China"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Jiade","family":"Huang","sequence":"additional","affiliation":[{"name":"Zhuzhou CRRC Times Electric Co., Ltd., Zhuzhou 412001, China"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Ming","family":"Gao","sequence":"additional","affiliation":[{"name":"State Key Laboratory of Advanced Design and Manufacturing for Vehicle Body, College of Mechanical and Vehicle Engineering, Hunan University, Changsha 410082, China"}],"role":[{"vocabulary":"crossref","role":"author"}]}],"member":"1968","published-online":{"date-parts":[[2023,6,27]]},"reference":[{"key":"ref_1","unstructured":"Balasubramaniam, A., and Pasricha, S. (2023). Object Detection in Autonomous Vehicles: Status and Open Challenges. arXiv."},{"key":"ref_2","doi-asserted-by":"crossref","unstructured":"Li, Y., Li, Z., Teng, S., Zhang, Y., Zhou, Y., Zhu, Y., Cao, D., Tian, B., Ai, Y., and Xuanyuan, Z. (2022, January 19\u201320). AutoMine: An Unmanned Mine Dataset. Proceedings of the 2022 IEEE\/CVF Conference on Computer Vision and Pattern Recognition (CVPR), New Orleans, LA, USA.","DOI":"10.1109\/CVPR52688.2022.02062"},{"key":"ref_3","doi-asserted-by":"crossref","unstructured":"Wang, C.-Y., Bochkovskiy, A., and Liao, H.-Y.M. (2023). YOLOv7: Trainable bag-of-freebies sets new state-of-the-art for real-time object detectors. arXiv.","DOI":"10.1109\/CVPR52729.2023.00721"},{"key":"ref_4","doi-asserted-by":"crossref","unstructured":"Ferrari, V., Hebert, M., Sminchisescu, C., and Weiss, Y. (2018, January 8\u201314). CornerNet: Detecting Objects as Paired Keypoints. Proceedings of the Computer Vision\u2014ECCV, Munich, Germany.","DOI":"10.1007\/978-3-030-01225-0"},{"key":"ref_5","doi-asserted-by":"crossref","unstructured":"Fleet, D., Pajdla, T., Schiele, B., and Tuytelaars, T. (2014, January 6\u201312). Microsoft COCO: Common Objects in Context. Proceedings of the Computer Vision\u2014ECCV 2014, Zurich, Switzerland.","DOI":"10.1007\/978-3-319-10599-4"},{"key":"ref_6","doi-asserted-by":"crossref","unstructured":"Cheng, G., Yuan, X., Yao, X., Yan, K., Zeng, Q., Xie, X., and Han, J. (2023). Towards Large-Scale Small Object Detection: Survey and Benchmarks. arXiv.","DOI":"10.1109\/TPAMI.2023.3290594"},{"key":"ref_7","doi-asserted-by":"crossref","unstructured":"Chen, C., Zhang, Y., Lv, Q., Wei, S., Wang, X., Sun, X., and Dong, J. (2019, January 27\u201328). RRNet: A Hybrid Detector for Object Detection in Drone-Captured Images. Proceedings of the 2019 IEEE\/CVF International Conference on Computer Vision Workshop (ICCVW), Seoul, Republic of Korea.","DOI":"10.1109\/ICCVW.2019.00018"},{"key":"ref_8","doi-asserted-by":"crossref","first-page":"1532","DOI":"10.1109\/TPAMI.2014.2300479","article-title":"Fast Feature Pyramids for Object Detection","volume":"36","author":"Appel","year":"2014","journal-title":"IEEE Trans. Pattern Anal. Mach. Intell."},{"key":"ref_9","doi-asserted-by":"crossref","unstructured":"Jiang, D., Sun, B., Su, S., Zuo, Z., Wu, P., and Tan, X. (2020). FASSD: A Feature Fusion and Spatial Attention-Based Single Shot Detector for Small Object Detection. Electronics, 9.","DOI":"10.3390\/electronics9091536"},{"key":"ref_10","doi-asserted-by":"crossref","unstructured":"Oliveira, G.L., Burgard, W., and Brox, T. (2016, January 9\u201314). Efficient deep models for monocular road segmentation. Proceedings of the 2016 IEEE\/RSJ International Conference on Intelligent Robots and Systems (IROS), Daejeon, Republic of Korea.","DOI":"10.1109\/IROS.2016.7759717"},{"key":"ref_11","doi-asserted-by":"crossref","first-page":"3448","DOI":"10.1109\/TITS.2022.3228042","article-title":"Deep Dual-Resolution Networks for Real-Time and Accurate Semantic Segmentation of Traffic Scenes","volume":"24","author":"Pan","year":"2023","journal-title":"IEEE Trans. Intell. Transp. Syst."},{"key":"ref_12","doi-asserted-by":"crossref","unstructured":"Asgarian, H., Amirkhani, A., and Shokouhi, S.B. (2021, January 28\u201329). Fast Drivable Area Detection for Autonomous Driving with Deep Learning. Proceedings of the 2021 5th International Conference on Pattern Recognition and Image Analysis (IPRIA), Kashan, Iran.","DOI":"10.1109\/IPRIA53572.2021.9483535"},{"key":"ref_13","doi-asserted-by":"crossref","unstructured":"Ren, L., Yang, C., Song, R., Chen, S., and Ai, Y. (2021, January 18\u201320). An Feature Fusion Object Detector for Autonomous Driving in Mining Area. Proceedings of the 2021 International Conference on Cyber-Physical Social Intelligence (ICCSI), Beijing, China.","DOI":"10.1109\/ICCSI53130.2021.9736214"},{"key":"ref_14","doi-asserted-by":"crossref","first-page":"2285","DOI":"10.1109\/TIV.2022.3221767","article-title":"MSFANet: A Light Weight Object Detector Based on Context Aggregation and Attention Mechanism for Autonomous Mining Truck","volume":"8","author":"Song","year":"2022","journal-title":"IEEE Trans. Intell. Veh."},{"key":"ref_15","doi-asserted-by":"crossref","unstructured":"Wei, Q., Song, R., Yang, X., and Ai, Y. (2022, January 8\u201312). A real-time semantic segmentation method for autonomous driving in surface mine. Proceedings of the 2022 IEEE 25th International Conference on Intelligent Transportation Systems (ITSC), Macau, China.","DOI":"10.1109\/ITSC55140.2022.9922492"},{"key":"ref_16","first-page":"1","article-title":"Open-Pit Mine Road Extraction From High-Resolution Remote Sensing Images Using RATT-UNet","volume":"19","author":"Xiao","year":"2022","journal-title":"IEEE Geosci. Remote Sens. Lett."},{"key":"ref_17","doi-asserted-by":"crossref","unstructured":"Lu, X., Ai, Y., and Tian, B. (2020). Real-Time Mine Road Boundary Detection and Tracking for Autonomous Truck. Sensors, 20.","DOI":"10.3390\/s20041121"},{"key":"ref_18","doi-asserted-by":"crossref","unstructured":"Kisantal, M., Wojna, Z., Murawski, J., Naruniec, J., and Cho, K. (2019). Augmentation for small object detection. arXiv.","DOI":"10.5121\/csit.2019.91713"},{"key":"ref_19","unstructured":"Wei, Z., Duan, C., Song, X., Tian, Y., and Wang, H. (2020). AMRNet: Chips Augmentation in Aerial Images Object Detection. arXiv."},{"key":"ref_20","doi-asserted-by":"crossref","unstructured":"Duan, C., Wei, Z., Zhang, C., Qu, S., and Wang, H. (2021, January 11\u201317). Coarse-grained Density Map Guided Object Detection in Aerial Images. Proceedings of the 2021 IEEE\/CVF International Conference on Computer Vision Workshops (ICCVW), Montreal, BC, Canada.","DOI":"10.1109\/ICCVW54120.2021.00313"},{"key":"ref_21","unstructured":"Bochkovskiy, A., Wang, C.-Y., and Liao, H.-Y.M. (2020). YOLOv4: Optimal Speed and Accuracy of Object Detection. arXiv."},{"key":"ref_22","doi-asserted-by":"crossref","unstructured":"Yang, F., Choi, W., and Lin, Y. (July, January 26). Exploit All the Layers: Fast and Accurate CNN Object Detector with Scale Dependent Pooling and Cascaded Rejection Classifiers. Proceedings of the 2016 IEEE Conference on Computer Vision and Pattern Recognition (CVPR), Las Vegas, Nevada.","DOI":"10.1109\/CVPR.2016.234"},{"key":"ref_23","unstructured":"Leibe, B., Matas, J., Sebe, N., and Welling, M. (2016, January 27\u201330). A Unified Multi-scale Deep Convolutional Neural Network for Fast Object Detection. Proceedings of the Computer Vision\u2014ECCV 2016, Amsterdam, The Netherlands."},{"key":"ref_24","doi-asserted-by":"crossref","first-page":"4701715","DOI":"10.1109\/TGRS.2021.3076050","article-title":"Oriented Bounding Boxes for Small and Freely Rotated Objects","volume":"60","author":"Zand","year":"2022","journal-title":"IEEE Trans. Geosci. Remote Sens."},{"key":"ref_25","doi-asserted-by":"crossref","unstructured":"Hu, J., Shen, L., Albanie, S., Sun, G., and Wu, E. (2018, January 8). Squeeze-and-Excitation Networks. Proceedings of the 2018 IEEE\/CVF Conference on Computer Vision and Pattern Recognition, Salt Lake City, UT, USA.","DOI":"10.1109\/CVPR.2018.00745"},{"key":"ref_26","doi-asserted-by":"crossref","unstructured":"Wang, Q., Wu, B., Zhu, P., Li, P., Zuo, W., and Hu, Q. (2020, January 13\u201319). ECA-Net: Efficient Channel Attention for Deep Convolutional Neural Networks. Proceedings of the 2020 IEEE\/CVF Conference on Computer Vision and Pattern Recognition (CVPR), Seattle, WA, USA.","DOI":"10.1109\/CVPR42600.2020.01155"},{"key":"ref_27","doi-asserted-by":"crossref","unstructured":"Ferrari, V., Hebert, M., Sminchisescu, C., and Weiss, Y. (2018, January 8). CBAM: Convolutional Block Attention Module. Proceedings of the Computer Vision\u2014ECCV, Munich, Germany.","DOI":"10.1007\/978-3-030-01216-8"},{"key":"ref_28","doi-asserted-by":"crossref","unstructured":"Long, J., Shelhamer, E., and Darrell, T. (2015, January 7\u201315). Fully convolutional networks for semantic segmentation. Proceedings of the 2015 IEEE Conference on Computer Vision and Pattern Recognition (CVPR), Boston, MA, USA.","DOI":"10.1109\/CVPR.2015.7298965"},{"key":"ref_29","doi-asserted-by":"crossref","unstructured":"Navab, N., Hornegger, J., Wells, W.M., and Frangi, A.F. (2015, January 5\u20139). U-Net: Convolutional Networks for Biomedical Image Segmentation. Proceedings of the Medical Image Computing and Computer-Assisted Intervention\u2013MICCAI, Munich, Germany.","DOI":"10.1007\/978-3-319-24553-9"},{"key":"ref_30","doi-asserted-by":"crossref","first-page":"834","DOI":"10.1109\/TPAMI.2017.2699184","article-title":"DeepLab: Semantic Image Segmentation with Deep Convolutional Nets, Atrous Convolution, and Fully Connected CRFs","volume":"40","author":"Chen","year":"2018","journal-title":"IEEE Trans. Pattern Anal. Mach. Intell."},{"key":"ref_31","unstructured":"Chen, L.-C., Papandreou, G., Schroff, F., and Adam, H. (2017). Rethinking Atrous Convolution for Semantic Image Segmentation. arXiv."},{"key":"ref_32","doi-asserted-by":"crossref","first-page":"550","DOI":"10.1007\/s11633-022-1339-y","article-title":"YOLOP: You Only Look Once for Panoptic Driving Perception","volume":"19","author":"Wu","year":"2022","journal-title":"Mach. Intell. Res."},{"key":"ref_33","unstructured":"Han, C., Zhao, Q., Zhang, S., Chen, Y., Zhang, Z., and Yuan, J. (2022). YOLOPv2: Better, Faster, Stronger for Panoptic Driving Perception. arXiv."},{"key":"ref_34","unstructured":"Vu, D., Ngo, B., and Phan, H. (2022). HybridNets: End-to-End Perception Network. arXiv."},{"key":"ref_35","doi-asserted-by":"crossref","unstructured":"Bolya, D., Zhou, C., Xiao, F., and Lee, Y.J. (November, January 27). YOLACT: Real-Time Instance Segmentation. Proceedings of the 2019 IEEE\/CVF International Conference on Computer Vision (ICCV), Seoul, Republic of Korea.","DOI":"10.1109\/ICCV.2019.00925"},{"key":"ref_36","doi-asserted-by":"crossref","unstructured":"He, K., Zhang, X., Ren, S., and Sun, J. (2016, January 27\u201330). Deep Residual Learning for Image Recognition. Proceedings of the 2016 IEEE Conference on Computer Vision and Pattern Recognition (CVPR), Las Vegas, NV, USA.","DOI":"10.1109\/CVPR.2016.90"},{"key":"ref_37","doi-asserted-by":"crossref","unstructured":"Huang, G., Liu, Z., Van Der Maaten, L., and Weinberger, K.Q. (2017, January 21\u201326). Densely Connected Convolutional Networks. Proceedings of the 2017 IEEE Conference on Computer Vision and Pattern Recognition (CVPR), Honolulu, HI, USA.","DOI":"10.1109\/CVPR.2017.243"},{"key":"ref_38","unstructured":"Gevorgyan, Z. (2022). SIoU Loss: More Powerful Learning for Bounding Box Regression. arXiv."}],"container-title":["Sensors"],"original-title":[],"language":"en","link":[{"URL":"https:\/\/www.mdpi.com\/1424-8220\/23\/13\/5977\/pdf","content-type":"unspecified","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2025,10,10]],"date-time":"2025-10-10T20:02:06Z","timestamp":1760126526000},"score":1,"resource":{"primary":{"URL":"https:\/\/www.mdpi.com\/1424-8220\/23\/13\/5977"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2023,6,27]]},"references-count":38,"journal-issue":{"issue":"13","published-online":{"date-parts":[[2023,7]]}},"alternative-id":["s23135977"],"URL":"https:\/\/doi.org\/10.3390\/s23135977","relation":{},"ISSN":["1424-8220"],"issn-type":[{"value":"1424-8220","type":"electronic"}],"subject":[],"published":{"date-parts":[[2023,6,27]]}}}