{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2026,7,16]],"date-time":"2026-07-16T12:37:24Z","timestamp":1784205444928,"version":"3.55.0"},"reference-count":33,"publisher":"MDPI AG","issue":"5","license":[{"start":{"date-parts":[[2024,5,13]],"date-time":"2024-05-13T00:00:00Z","timestamp":1715558400000},"content-version":"vor","delay-in-days":0,"URL":"https:\/\/creativecommons.org\/licenses\/by\/4.0\/"}],"funder":[{"DOI":"10.13039\/501100001809","name":"National Natural Science Foundation of China","doi-asserted-by":"publisher","award":["61573286"],"award-info":[{"award-number":["61573286"]}],"id":[{"id":"10.13039\/501100001809","id-type":"DOI","asserted-by":"publisher"}]},{"DOI":"10.13039\/501100001809","name":"National Natural Science Foundation of China","doi-asserted-by":"publisher","award":["201905053003"],"award-info":[{"award-number":["201905053003"]}],"id":[{"id":"10.13039\/501100001809","id-type":"DOI","asserted-by":"publisher"}]},{"name":"Aeronautical Science Foundation of China","award":["61573286"],"award-info":[{"award-number":["61573286"]}]},{"name":"Aeronautical Science Foundation of China","award":["201905053003"],"award-info":[{"award-number":["201905053003"]}]}],"content-domain":{"domain":[],"crossmark-restriction":false},"short-container-title":["IJGI"],"abstract":"<jats:p>A deep learning-based Visual Inertial SLAM technique is proposed in this paper to ensure accurate autonomous localization of mobile robots in environments with dynamic objects. Addressing the limitations of real-time performance in deep learning algorithms and the poor robustness of pure visual geometry algorithms, this paper presents a deep learning-based Visual Inertial SLAM technique. Firstly, a non-blocking model is designed to extract semantic information from images. Then, a motion probability hierarchy model is proposed to obtain prior motion probabilities of feature points. For image frames without semantic information, a motion probability propagation model is designed to determine the prior motion probabilities of feature points. Furthermore, considering that the output of inertial measurements is unaffected by dynamic objects, this paper integrates inertial measurement information to improve the estimation accuracy of feature point motion probabilities. An adaptive threshold-based motion probability estimation method is proposed, and finally, the positioning accuracy is enhanced by eliminating feature points with excessively high motion probabilities. Experimental results demonstrate that the proposed algorithm achieves accurate localization in dynamic environments while maintaining real-time performance.<\/jats:p>","DOI":"10.3390\/ijgi13050163","type":"journal-article","created":{"date-parts":[[2024,5,13]],"date-time":"2024-05-13T08:33:03Z","timestamp":1715589183000},"page":"163","update-policy":"https:\/\/doi.org\/10.3390\/mdpi_crossmark_policy","source":"Crossref","is-referenced-by-count":16,"title":["VIS-SLAM: A Real-Time Dynamic SLAM Algorithm Based on the Fusion of Visual, Inertial, and Semantic Information"],"prefix":"10.3390","volume":"13","author":[{"given":"Yinglong","family":"Wang","sequence":"first","affiliation":[{"name":"School of Automation, Northwestern Polytechnical University, Xi\u2019an 710072, China"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"ORCID":"https:\/\/orcid.org\/0000-0002-8061-5974","authenticated-orcid":false,"given":"Xiaoxiong","family":"Liu","sequence":"additional","affiliation":[{"name":"Shaanxi Province Key Laboratory of Flight Control and Simulation Technology, Xi\u2019an 710072, China"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Minkun","family":"Zhao","sequence":"additional","affiliation":[{"name":"School of Automation, Northwestern Polytechnical University, Xi\u2019an 710072, China"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Xinlong","family":"Xu","sequence":"additional","affiliation":[{"name":"School of Automation, Northwestern Polytechnical University, Xi\u2019an 710072, China"}],"role":[{"vocabulary":"crossref","role":"author"}]}],"member":"1968","published-online":{"date-parts":[[2024,5,13]]},"reference":[{"key":"ref_1","doi-asserted-by":"crossref","unstructured":"Lu, X., Wang, H., Tang, S., Huang, H., and Li, C. (2020). DM-SLAM: Monocular SLAM in dynamic environments. Appl. Sci., 10.","DOI":"10.20944\/preprints202001.0123.v1"},{"key":"ref_2","doi-asserted-by":"crossref","first-page":"110","DOI":"10.1016\/j.robot.2016.11.012","article-title":"Improving RGB-D SLAM in dynamic environments: A motion removal approach","volume":"89","author":"Sun","year":"2017","journal-title":"Robot. Auton. Syst."},{"key":"ref_3","unstructured":"Chum, O., and Matas, J. (2005, January 20\u201326). Matching with PROSAC-progressive sample consensus. Proceedings of the 2005 IEEE Computer Society Conference on Computer Vision and Pattern Recognition (CVPR\u201905), San Diego, CA, USA."},{"key":"ref_4","doi-asserted-by":"crossref","first-page":"1004","DOI":"10.1109\/TRO.2018.2853729","article-title":"Vins-mono: A robust and versatile monocular visual-inertial state estimator","volume":"34","author":"Qin","year":"2018","journal-title":"IEEE Trans. Robot."},{"key":"ref_5","doi-asserted-by":"crossref","first-page":"115","DOI":"10.1016\/j.robot.2018.07.002","article-title":"Motion removal for reliable RGB-D SLAM in dynamic environments","volume":"108","author":"Sun","year":"2018","journal-title":"Robot. Auton. Syst."},{"key":"ref_6","doi-asserted-by":"crossref","unstructured":"Wang, R., Wan, W., Wang, Y., and Di, K. (2019). A new RGB-D SLAM method with moving object detection for dynamic indoor scenes. Remote Sens., 11.","DOI":"10.3390\/rs11101143"},{"key":"ref_7","doi-asserted-by":"crossref","unstructured":"Zhang, C., Zhang, R., Jin, S., and Yi, X. (2022). PFD-SLAM: A new RGB-D SLAM for dynamic indoor environments based on non-prior semantic segmentation. Remote Sens., 14.","DOI":"10.3390\/rs14102445"},{"key":"ref_8","doi-asserted-by":"crossref","first-page":"373","DOI":"10.1109\/TPAMI.2020.3010942","article-title":"Rgb-d slam in dynamic environments using point correlations","volume":"44","author":"Dai","year":"2020","journal-title":"IEEE Trans. Pattern Anal. Mach. Intell."},{"key":"ref_9","doi-asserted-by":"crossref","first-page":"469","DOI":"10.1145\/235815.235821","article-title":"The quickhull algorithm for convex hulls","volume":"22","author":"Barber","year":"1996","journal-title":"ACM Trans. Math. Softw. (TOMS)"},{"key":"ref_10","doi-asserted-by":"crossref","unstructured":"Yu, C., Liu, Z., Liu, X.J., Xie, F., Yang, Y., Wei, Q., and Fei, Q. (2018, January 1\u20135). DS-SLAM: A semantic visual SLAM towards dynamic environments. Proceedings of the 2018 IEEE\/RSJ International Conference on Intelligent Robots and Systems (IROS), Madrid, Spain.","DOI":"10.1109\/IROS.2018.8593691"},{"key":"ref_11","doi-asserted-by":"crossref","first-page":"1255","DOI":"10.1109\/TRO.2017.2705103","article-title":"Orb-slam2: An open-source slam system for monocular, stereo, and rgb-d cameras","volume":"33","year":"2017","journal-title":"IEEE Trans. Robot."},{"key":"ref_12","doi-asserted-by":"crossref","first-page":"2481","DOI":"10.1109\/TPAMI.2016.2644615","article-title":"Segnet: A deep convolutional encoder-decoder architecture for image segmentation","volume":"39","author":"Badrinarayanan","year":"2017","journal-title":"IEEE Trans. Pattern Anal. Mach. Intell."},{"key":"ref_13","doi-asserted-by":"crossref","first-page":"4076","DOI":"10.1109\/LRA.2018.2860039","article-title":"DynaSLAM: Tracking, mapping, and inpainting in dynamic scenes","volume":"3","author":"Bescos","year":"2018","journal-title":"IEEE Robot. Autom. Lett."},{"key":"ref_14","doi-asserted-by":"crossref","first-page":"23772","DOI":"10.1109\/ACCESS.2021.3050617","article-title":"RDS-SLAM: Real-time dynamic SLAM using semantic segmentation methods","volume":"9","author":"Liu","year":"2021","journal-title":"IEEE Access"},{"key":"ref_15","doi-asserted-by":"crossref","first-page":"155047","DOI":"10.1109\/ACCESS.2020.3018557","article-title":"Real-time visual-inertial localization using semantic segmentation towards dynamic environments","volume":"8","author":"Zhao","year":"2020","journal-title":"IEEE Access"},{"key":"ref_16","doi-asserted-by":"crossref","unstructured":"Yurtkulu, S.C., \u015eahin, Y.H., and Unal, G. (2019, January 24\u201326). Semantic segmentation with extended DeepLabv3 architecture. Proceedings of the 2019 27th Signal Processing and Communications Applications Conference (SIU), Sivas, Turkey.","DOI":"10.1109\/SIU.2019.8806244"},{"key":"ref_17","doi-asserted-by":"crossref","first-page":"11523","DOI":"10.1109\/LRA.2022.3203231","article-title":"DynaVINS: A visual-inertial SLAM for dynamic environments","volume":"7","author":"Song","year":"2022","journal-title":"IEEE Robot. Autom. Lett."},{"key":"ref_18","doi-asserted-by":"crossref","first-page":"9573","DOI":"10.1109\/LRA.2022.3191193","article-title":"RGB-D inertial odometry for a resource-restricted robot in dynamic environments","volume":"7","author":"Liu","year":"2022","journal-title":"IEEE Robot. Autom. Lett."},{"key":"ref_19","doi-asserted-by":"crossref","unstructured":"Zhu, X., Lyu, S., Wang, X., and Zhao, Q. (2021, January 11\u201317). TPH-YOLOv5: Improved YOLOv5 based on transformer prediction head for object detection on drone-captured scenarios. Proceedings of the IEEE\/CVF International Conference on Computer Vision, Virtual.","DOI":"10.1109\/ICCVW54120.2021.00312"},{"key":"ref_20","doi-asserted-by":"crossref","unstructured":"Sun, Y., Wang, Q., Yan, C., Feng, Y., Tan, R., Shi, X., and Wang, X. (2023). D-VINS: Dynamic adaptive visual\u2013inertial SLAM with IMU prior and semantic constraints in dynamic scenes. Remote Sens., 15.","DOI":"10.20944\/preprints202305.2154.v1"},{"key":"ref_21","doi-asserted-by":"crossref","first-page":"20939","DOI":"10.1007\/s00521-023-08809-1","article-title":"An improved fire detection approach based on YOLO-v8 for smart cities","volume":"35","author":"Talaat","year":"2023","journal-title":"Neural Comput. Appl."},{"key":"ref_22","doi-asserted-by":"crossref","first-page":"1","DOI":"10.1016\/j.robot.2019.03.012","article-title":"Dynamic-SLAM: Semantic monocular visual localization and mapping based on deep learning in dynamic environment","volume":"117","author":"Xiao","year":"2019","journal-title":"Robot. Auton. Syst."},{"key":"ref_23","unstructured":"Veit, A., Matera, T., Neumann, L., Matas, J., and Belongie, S. (2016). Coco-text: Dataset and benchmark for text detection and recognition in natural images. arXiv."},{"key":"ref_24","doi-asserted-by":"crossref","first-page":"573","DOI":"10.1037\/a0029146","article-title":"Bayesian estimation supersedes the t test","volume":"142","author":"Kruschke","year":"2013","journal-title":"J. Exp. Psychol. Gen."},{"key":"ref_25","doi-asserted-by":"crossref","first-page":"1","DOI":"10.1109\/TRO.2016.2597321","article-title":"On-manifold preintegration for real-time visual\u2013inertial odometry","volume":"33","author":"Forster","year":"2016","journal-title":"IEEE Trans. Robot."},{"key":"ref_26","doi-asserted-by":"crossref","unstructured":"Shen, S., Michael, N., and Kumar, V. (2015, January 26\u201330). Tightly-coupled monocular visual-inertial fusion for autonomous flight of rotorcraft MAVs. Proceedings of the 2015 IEEE International Conference on Robotics and Automation (ICRA), Seattle, WA, USA.","DOI":"10.1109\/ICRA.2015.7139939"},{"key":"ref_27","doi-asserted-by":"crossref","unstructured":"Miao, S., Liu, X., Wei, D., and Li, C. (2021). A visual SLAM robust against dynamic objects based on hybrid semantic-geometry information. ISPRS Int. J. Geo-Inf., 10.","DOI":"10.3390\/ijgi10100673"},{"key":"ref_28","doi-asserted-by":"crossref","unstructured":"Triggs, B., McLauchlan, P.F., Hartley, R.I., and Fitzgibbon, A.W. (1999, January 21\u201322). Bundle adjustment\u2014A modern synthesis. Proceedings of the Vision Algorithms: Theory and Practice: International Workshop on Vision Algorithms, Corfu, Greece.","DOI":"10.1007\/3-540-44480-7_21"},{"key":"ref_29","doi-asserted-by":"crossref","first-page":"5393","DOI":"10.30534\/ijatcse\/2020\/175942020","article-title":"Binary cross entropy with deep learning technique for image classification","volume":"9","author":"Ruby","year":"2020","journal-title":"Int. J. Adv. Trends Comput. Sci. Eng."},{"key":"ref_30","doi-asserted-by":"crossref","unstructured":"She, Q., Feng, F., Hao, X., Yang, Q., Lan, C., Lomonaco, V., Shi, X., Wang, Z., Guo, Y., and Zhang, Y. (August, January 31). Openloris-object: A robotic vision dataset and benchmark for lifelong deep learning. Proceedings of the 2020 IEEE International Conference on Robotics and Automation (ICRA), Paris, France.","DOI":"10.1109\/ICRA40945.2020.9196887"},{"key":"ref_31","doi-asserted-by":"crossref","first-page":"1874","DOI":"10.1109\/TRO.2021.3075644","article-title":"Orb-slam3: An accurate open-source library for visual, visual\u2013inertial, and multimap slam","volume":"37","author":"Campos","year":"2021","journal-title":"IEEE Trans. Robot."},{"key":"ref_32","doi-asserted-by":"crossref","unstructured":"Sturm, J., Engelhard, N., Endres, F., Burgard, W., and Cremers, D. (2012, January 7\u201312). A benchmark for the evaluation of RGB-D SLAM systems. Proceedings of the 2012 IEEE\/RSJ International Conference on Intelligent Robots and Systems, Algarve, Portugal.","DOI":"10.1109\/IROS.2012.6385773"},{"key":"ref_33","doi-asserted-by":"crossref","unstructured":"Quigley, M., Conley, K., Gerkey, B., Faust, J., Foote, T., Leibs, J., Wheeler, R., and Ng, A.Y. (2009, January 12\u201317). ROS: An open-source Robot Operating System. Proceedings of the ICRA Workshop on Open Source Software, Kobe, Japan.","DOI":"10.1109\/MRA.2010.936956"}],"container-title":["ISPRS International Journal of Geo-Information"],"original-title":[],"language":"en","link":[{"URL":"https:\/\/www.mdpi.com\/2220-9964\/13\/5\/163\/pdf","content-type":"unspecified","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2025,10,10]],"date-time":"2025-10-10T14:41:27Z","timestamp":1760107287000},"score":1,"resource":{"primary":{"URL":"https:\/\/www.mdpi.com\/2220-9964\/13\/5\/163"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2024,5,13]]},"references-count":33,"journal-issue":{"issue":"5","published-online":{"date-parts":[[2024,5]]}},"alternative-id":["ijgi13050163"],"URL":"https:\/\/doi.org\/10.3390\/ijgi13050163","relation":{},"ISSN":["2220-9964"],"issn-type":[{"value":"2220-9964","type":"electronic"}],"subject":[],"published":{"date-parts":[[2024,5,13]]}}}