{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2026,7,24]],"date-time":"2026-07-24T15:16:48Z","timestamp":1784906208507,"version":"3.55.0"},"reference-count":68,"publisher":"Springer Science and Business Media LLC","issue":"3","license":[{"start":{"date-parts":[[2025,9,15]],"date-time":"2025-09-15T00:00:00Z","timestamp":1757894400000},"content-version":"tdm","delay-in-days":0,"URL":"https:\/\/creativecommons.org\/licenses\/by\/4.0"},{"start":{"date-parts":[[2025,9,15]],"date-time":"2025-09-15T00:00:00Z","timestamp":1757894400000},"content-version":"vor","delay-in-days":0,"URL":"https:\/\/creativecommons.org\/licenses\/by\/4.0"}],"funder":[{"name":"Leonardo UK"},{"DOI":"10.13039\/501100000266","name":"Engineering and Physical Sciences Research Council","doi-asserted-by":"publisher","award":["EP\/V026763\/1"],"award-info":[{"award-number":["EP\/V026763\/1"]}],"id":[{"id":"10.13039\/501100000266","id-type":"DOI","asserted-by":"publisher"}]},{"DOI":"10.13039\/501100000266","name":"Engineering and Physical Sciences Research Council","doi-asserted-by":"publisher","award":["EP\/V026763\/1"],"award-info":[{"award-number":["EP\/V026763\/1"]}],"id":[{"id":"10.13039\/501100000266","id-type":"DOI","asserted-by":"publisher"}]}],"content-domain":{"domain":["link.springer.com"],"crossmark-restriction":false},"short-container-title":["J Intell Robot Syst"],"abstract":"<jats:title>Abstract<\/jats:title>\n          <jats:p>Deep Reinforcement learning (DRL) is used to enable autonomous navigation in unknown environments. Most research assumes perfect sensor data, but real-world environments may contain natural and artificial sensor noise and denial. Here, we present a benchmark of both well-used and emerging DRL algorithms in two navigation tasks - Lidar + position, and vision end-to-end - with configurable sensor denial effects. In particular, we are interested in comparing how different DRL methods (e.g. model-free, on-policy PPO vs. model-free off-policy TD3, vs. model-based DreamerV3) are affected by imperfect sensor readings. We show that DreamerV3 outperforms other methods in the visual end-to-end navigation task with a dynamic goal. Furthermore, DreamerV3 generally outperforms other methods in sensor-denied environments. In order to improve robustness, we use adversarial training and demonstrate an improved performance in denied environments, although we show that this may lead to the agent learning to choose high-risk actions in case of uncertain sensor readings, which is not appropriate for safety-critical scenarios. We anticipate this benchmark of different DRL methods and the usage of adversarial training to be a starting point for the development of more elaborate navigation strategies that are capable of dealing with uncertain and denied sensor readings.<\/jats:p>","DOI":"10.1007\/s10846-025-02309-1","type":"journal-article","created":{"date-parts":[[2025,9,15]],"date-time":"2025-09-15T07:20:45Z","timestamp":1757920845000},"update-policy":"https:\/\/doi.org\/10.1007\/springer_crossmark_policy","source":"Crossref","is-referenced-by-count":8,"title":["Benchmarking Deep Reinforcement Learning for Navigation in Denied Sensor Environments"],"prefix":"10.1007","volume":"111","author":[{"ORCID":"https:\/\/orcid.org\/0000-0001-6795-6216","authenticated-orcid":false,"given":"Mariusz","family":"Wisniewski","sequence":"first","affiliation":[],"role":[{"vocabulary":"crossref","role":"author"}]},{"ORCID":"https:\/\/orcid.org\/0009-0006-3212-0578","authenticated-orcid":false,"given":"Paraskevas","family":"Chatzithanos","sequence":"additional","affiliation":[],"role":[{"vocabulary":"crossref","role":"author"}]},{"ORCID":"https:\/\/orcid.org\/0000-0003-3524-3953","authenticated-orcid":false,"given":"Weisi","family":"Guo","sequence":"additional","affiliation":[],"role":[{"vocabulary":"crossref","role":"author"}]},{"ORCID":"https:\/\/orcid.org\/0000-0002-3966-7633","authenticated-orcid":false,"given":"Antonios","family":"Tsourdos","sequence":"additional","affiliation":[],"role":[{"vocabulary":"crossref","role":"author"}]}],"member":"297","published-online":{"date-parts":[[2025,9,15]]},"reference":[{"key":"2309_CR1","doi-asserted-by":"publisher","unstructured":"Krizhevsky, A., Sutskever, I., Hinton, G.E.: ImageNet classification with deep convolutional neural networks. Commun. ACM 60(6), 84\u201390 (2017). https:\/\/doi.org\/10.1145\/3065386. Accessed 2021-07-21","DOI":"10.1145\/3065386"},{"key":"2309_CR2","doi-asserted-by":"publisher","unstructured":"Deng, J., Dong, W., Socher, R., Li, L.-J., Kai Li, Li Fei-Fei: ImageNet: A large-scale hierarchical image database. In: 2009 IEEE Conference on Computer Vision and Pattern Recognition, pp. 248\u2013255. IEEE, Miami, FL (2009). https:\/\/doi.org\/10.1109\/CVPR.2009.5206848. https:\/\/ieeexplore.ieee.org\/document\/5206848\/ Accessed 2021-07-21","DOI":"10.1109\/CVPR.2009.5206848"},{"key":"2309_CR3","unstructured":"He, K., Zhang, X., Ren, S., Sun, J.: Deep Residual Learning for Image Recognition. arXiv:1512.03385 [cs] (2015). Accessed 2021-07-10"},{"key":"2309_CR4","doi-asserted-by":"publisher","unstructured":"Mnih, V., Kavukcuoglu, K., Silver, D., Rusu, A.A., Veness, J., Bellemare, M.G., Graves, A., Riedmiller, M., Fidjeland, A.K., Ostrovski, G., Petersen, S., Beattie, C., Sadik, A., Antonoglou, I., King, H., Kumaran, D., Wierstra, D., Legg, S., Hassabis, D.: Human-level control through deep reinforcement learning. Nature 518(7540), 529\u2013533 (2015). https:\/\/doi.org\/10.1038\/nature14236. Accessed 2022-12-10","DOI":"10.1038\/nature14236"},{"key":"2309_CR5","doi-asserted-by":"publisher","unstructured":"Geles, I., Bauersfeld, L., Romero, A., Xing, J., Scaramuzza, D.: Demonstrating Agile Flight from Pixels without State Estimation. In: Robotics: Science and Systems XX. Robotics: Science and Systems Foundation, ??? (2024). https:\/\/doi.org\/10.15607\/RSS.2024.XX.082. http:\/\/www.roboticsproceedings.org\/rss20\/p082.pdf Accessed 2025-03-31","DOI":"10.15607\/RSS.2024.XX.082"},{"key":"2309_CR6","doi-asserted-by":"publisher","unstructured":"Walker, O., Vanegas, F., Gonzalez, F., Koenig, S.: A deep reinforcement learning framework for uav navigation in indoor environments. In: IEEE Aerospace Conference, pp. 1\u201314 (2019). https:\/\/doi.org\/10.1109\/AERO.2019.8742226","DOI":"10.1109\/AERO.2019.8742226"},{"key":"2309_CR7","unstructured":"Zhang, H., Chen, H., Xiao, C., Li, B., Liu, M., Boning, D., Hsieh, C.-J.: Robust Deep Reinforcement Learning against Adversarial Perturbations on State Observations. In: Advances in Neural Information Processing Systems, 33, pp. 21024\u201321037. Curran Associates, Inc., ??? (2020). https:\/\/proceedings.neurips.cc\/paper_files\/paper\/2020\/hash\/f0eb6568ea114ba6e293f903c34d7488-Abstract.html Accessed 2024-10-09"},{"key":"2309_CR8","unstructured":"Mirowski, P., Pascanu, R., Viola, F., Soyer, H., Ballard, A.J., Banino, A., Denil, M., Goroshin, R., Sifre, L., Kavukcuoglu, K., Kumaran, D., Hadsell, R.: Learning to Navigate in Complex Environments. arXiv. arXiv:1611.03673 [cs] (2017). Accessed 2022-10-17"},{"key":"2309_CR9","unstructured":"Cimurs, R., Suh, I.H., Lee, J.H.: Goal-Driven Autonomous Exploration Through Deep Reinforcement Learning. arXiv. arXiv:2103.07119 [cs] (2021). Accessed 2023-10-19"},{"key":"2309_CR10","doi-asserted-by":"publisher","unstructured":"Wisniewski, M., Ona, I.G., Chatzithanos, P., Guo, W., Tsourdos, A.: Autonomous Navigation in Dynamic Maze Environments Under Adversarial Sensor Attack. In: 2024 International Conference on Unmanned Aircraft Systems (ICUAS), pp. 881\u2013886 (2024). https:\/\/doi.org\/10.1109\/ICUAS60882.2024.10557088 . ISSN: 2575-7296. https:\/\/ieeexplore.ieee.org\/document\/10557088\/?arnumber=10557088 Accessed 2024-08-23","DOI":"10.1109\/ICUAS60882.2024.10557088"},{"key":"2309_CR11","doi-asserted-by":"publisher","unstructured":"Zhu, K., Zhang, T.: Deep reinforcement learning based mobile robot navigation: A review. Tsinghua Science and Technology 26(5), 674\u2013691 (2021) https:\/\/doi.org\/10.26599\/TST.2021.9010012. Conference Name: Tsinghua Science and Technology. Accessed 2024-08-27","DOI":"10.26599\/TST.2021.9010012"},{"key":"2309_CR12","unstructured":"Beattie, C., Leibo, J.Z., Teplyashin, D., Ward, T., Wainwright, M., K\u00fcttler, H., Lefrancq, A., Green, S., Vald\u00e9s, V., Sadik, A., Schrittwieser, J., Anderson, K., York, S., Cant, M., Cain, A., Bolton, A., Gaffney, S., King, H., Hassabis, D., Legg, S., Petersen, S.: DeepMind Lab. arXiv. arXiv:1612.03801 [cs] (2016). Accessed 2023-05-19"},{"key":"2309_CR13","doi-asserted-by":"crossref","unstructured":"Kempka, M., Wydmuch, M., Runc, G., Toczek, J., Ja\u015bkowski, W.: ViZDoom: A Doom-based AI Research Platform for Visual Reinforcement Learning. arXiv:1605.02097 [cs] (2016). Accessed 2022-12-10","DOI":"10.1109\/CIG.2016.7860433"},{"key":"2309_CR14","unstructured":"Mnih, V., Badia, A.P., Mirza, M., Graves, A., Lillicrap, T.P., Harley, T., Silver, D., Kavukcuoglu, K.: Asynchronous Methods for Deep Reinforcement Learning. arXiv:1602.01783 [cs] (2016). Accessed 2022-10-24"},{"key":"2309_CR15","doi-asserted-by":"publisher","unstructured":"Kaufmann, E., Bauersfeld, L., Loquercio, A., M\u00fcller, M., Koltun, V., Scaramuzza, D.: Champion-level drone racing using deep reinforcement learning. Nature 620(7976), 982\u2013987 (2023). https:\/\/doi.org\/10.1038\/s41586-023-06419-4. Number: 7976 Publisher: Nature Publishing Group. Accessed 2023-09-10","DOI":"10.1038\/s41586-023-06419-4"},{"key":"2309_CR16","doi-asserted-by":"publisher","unstructured":"Polvara, R., Patacchiola, M., Sharma, S., Wan, J., Manning, A., Sutton, R., Cangelosi, A.: Toward end-to-end control for uav autonomous landing via deep reinforcement learning. In: International Conference on Unmanned Aircraft Systems (ICUAS), pp. 115\u2013123 (2018). https:\/\/doi.org\/10.1109\/ICUAS.2018.8453449","DOI":"10.1109\/ICUAS.2018.8453449"},{"issue":"6","key":"2309_CR17","doi-asserted-by":"publisher","first-page":"5068","DOI":"10.1109\/TITS.2020.3046646","volume":"23","author":"J Chen","year":"2022","unstructured":"Chen, J., Li, S.E., Tomizuka, M.: Interpretable end-to-end urban autonomous driving with latent deep reinforcement learning. IEEE Trans. Intell. Transp. Syst. 23(6), 5068\u20135078 (2022). https:\/\/doi.org\/10.1109\/TITS.2020.3046646","journal-title":"IEEE Trans. Intell. Transp. Syst."},{"key":"2309_CR18","unstructured":"Pasukonis, J., Lillicrap, T., Hafner, D.: Evaluating Long-Term Memory in 3D Mazes. arXiv:2210.13383 [cs] (2022). Accessed 2024-04-24"},{"key":"2309_CR19","unstructured":"Lample, G., Chaplot, D.S.: Playing FPS Games with Deep Reinforcement Learning. arXiv:1609.05521 [cs] (2018). Accessed 2022-10-14"},{"issue":"7","key":"2309_CR20","doi-asserted-by":"publisher","first-page":"1279","DOI":"10.1109\/TNN.2008.2000394","volume":"19","author":"C Luo","year":"2008","unstructured":"Luo, C., Yang, S.X.: A bioinspired neural network for real-time concurrent map building and complete coverage robot navigation in unknown environments. IEEE Trans. Neural Netw. 19(7), 1279\u20131298 (2008). https:\/\/doi.org\/10.1109\/TNN.2008.2000394","journal-title":"IEEE Trans. Neural Netw."},{"key":"2309_CR21","doi-asserted-by":"publisher","unstructured":"Chahine, M., Hasani, R., Kao, P., Ray, A., Shubert, R., Lechner, M., Amini, A., Rus, D.: Robust flight navigation out of distribution with liquid neural networks. Sci. Robot. 8(77), 8892 (2023). https:\/\/doi.org\/10.1126\/scirobotics.adc8892https:\/\/www.science.org\/doi\/pdf\/10.1126\/scirobotics.adc8892","DOI":"10.1126\/scirobotics.adc8892"},{"key":"2309_CR22","doi-asserted-by":"publisher","unstructured":"Sajjadi, S., Bittick, J., Janabi-Sharifi, F., Mantegh, I.: A robust and adaptive sensor fusion approach for indoor uav localization. In: 2023 International Conference on Unmanned Aircraft Systems (ICUAS), pp. 441\u2013447 (2023). https:\/\/doi.org\/10.1109\/ICUAS57906.2023.10156526","DOI":"10.1109\/ICUAS57906.2023.10156526"},{"key":"2309_CR23","doi-asserted-by":"publisher","unstructured":"Negru, S.A., Geragersian, P., Petrunin, I., Guo, W.: Resilient multi-sensor uav navigation with a hybrid federated fusion architecture. Sensors 24(3) (2024). https:\/\/doi.org\/10.3390\/s24030981","DOI":"10.3390\/s24030981"},{"issue":"1","key":"2309_CR24","doi-asserted-by":"publisher","first-page":"159","DOI":"10.1109\/TITS.2018.2889923","volume":"21","author":"C Liu","year":"2020","unstructured":"Liu, C., Zhang, G., Guo, W., He, R.: Kalman prediction-based neighbor discovery and its effect on routing protocol in vehicular ad hoc networks. IEEE Trans. Intell. Transp. Syst. 21(1), 159\u2013169 (2020). https:\/\/doi.org\/10.1109\/TITS.2018.2889923","journal-title":"IEEE Trans. Intell. Transp. Syst."},{"key":"2309_CR25","doi-asserted-by":"publisher","unstructured":"Bae, S., Han, D., Park, S.: A new approach to lidar and camera fusion for autonomous driving. In: 2023 International Conference on Artificial Intelligence in Information and Communication (ICAIIC), pp. 751\u2013753 (2023). https:\/\/doi.org\/10.1109\/ICAIIC57133.2023.10066963","DOI":"10.1109\/ICAIIC57133.2023.10066963"},{"key":"2309_CR26","doi-asserted-by":"publisher","unstructured":"Kacker, T., Perrusquia, A., Guo, W.: Multi-spectral fusion using generative adversarial networks for uav detection of wild fires. In: 2023 International Conference on Artificial Intelligence in Information and Communication (ICAIIC), pp. 182\u2013187 (2023). https:\/\/doi.org\/10.1109\/ICAIIC57133.2023.10067042","DOI":"10.1109\/ICAIIC57133.2023.10067042"},{"key":"2309_CR27","doi-asserted-by":"publisher","unstructured":"Spaan, M.T.J.: Partially Observable Markov Decision Processes. In: Wiering, M., Otterlo, M. (eds.) Reinforcement Learning: State-of-the-Art, pp. 387\u2013414. Springer, Berlin, Heidelberg (2012). https:\/\/doi.org\/10.1007\/978-3-642-27645-3_12 Accessed 2024-10-12","DOI":"10.1007\/978-3-642-27645-3_12"},{"key":"2309_CR28","unstructured":"Osogami, T.: Robust partially observable Markov decision process. In: Proceedings of the 32nd International Conference on Machine Learning, pp. 106\u2013115. PMLR, ??? (2015). ISSN: 1938-7228. https:\/\/proceedings.mlr.press\/v37\/osogami15.html Accessed 2024-10-12"},{"key":"2309_CR29","unstructured":"Tessler, C., Efroni, Y., Mannor, S.: Action Robust Reinforcement Learning and Applications in Continuous Control. In: Proceedings of the 36th International Conference on Machine Learning, pp. 6215\u20136224. PMLR, ??? (2019). ISSN: 2640-3498. https:\/\/proceedings.mlr.press\/v97\/tessler19a.html Accessed 2024-03-04"},{"key":"2309_CR30","unstructured":"Pinto, L., Davidson, J., Sukthankar, R., Gupta, A.: Robust Adversarial Reinforcement Learning. In: Proceedings of the 34th International Conference on Machine Learning, pp. 2817\u20132826. PMLR, ??? (2017). ISSN: 2640-3498. https:\/\/proceedings.mlr.press\/v70\/pinto17a.html Accessed 2025-05-28"},{"key":"2309_CR31","unstructured":"Korkmaz, E.: Adversarial Training Blocks Generalization in Neural Policies. (2021). https:\/\/openreview.net\/forum?id=fXGimmbtD9c Accessed 2025-03-26"},{"key":"2309_CR32","doi-asserted-by":"publisher","unstructured":"Korkmaz, E.: Adversarial Robust Deep Reinforcement Learning Requires Redefining Robustness. Proceed. AAAI Conf. Artif. Intell. 37(7), 8369\u20138377 (2023). https:\/\/doi.org\/10.1609\/aaai.v37i7.26009. Accessed 2024-09-28","DOI":"10.1609\/aaai.v37i7.26009"},{"key":"2309_CR33","unstructured":"Korkmaz, E.: Understanding and Diagnosing Deep Reinforcement Learning. In: Proceedings of the 41st International Conference on Machine Learning, pp. 25265\u201325275. PMLR, ??? (2024). ISSN: 2640-3498. https:\/\/proceedings.mlr.press\/v235\/korkmaz24a.html Accessed 2025-06-23"},{"key":"2309_CR34","unstructured":"Havens, A.J., Jiang, Z., Sarkar, S.: Online Robust Policy Learning in the Presence of Unknown Adversaries. arXiv:1807.06064 [cs, stat] (2018). Accessed 2024-10-09"},{"key":"2309_CR35","doi-asserted-by":"publisher","unstructured":"Panda, D.K., Guo, W.: Action Robust Reinforcement Learning for Air Mobility Deconfliction Against Conflict Induced Spoofing. IEEE Trans. Intell. Transp. Syst. 25(12), 21343\u201321355 (2024). https:\/\/doi.org\/10.1109\/TITS.2024.3454354. Accessed 2025-06-30","DOI":"10.1109\/TITS.2024.3454354"},{"key":"2309_CR36","doi-asserted-by":"publisher","unstructured":"Peng, Z., Song, X., Song, S., Stojanovic, V.: Spatiotemporal fault estimation for switched nonlinear reaction\u2013diffusion systems via adaptive iterative learning. Int. J. Adapt. Contr. Signal Process. 38(10), 3473\u20133483 (2024) https:\/\/doi.org\/10.1002\/acs.3885. _eprint: https:\/\/onlinelibrary.wiley.com\/doi\/pdf\/10.1002\/acs.3885. Accessed 2025-03-26","DOI":"10.1002\/acs.3885"},{"key":"2309_CR37","doi-asserted-by":"publisher","unstructured":"Song, X., Wu, C., Song, S., Stojanovic, V., Tejado, I.: Fuzzy wavelet neural adaptive finite-time self-triggered fault-tolerant control for a quadrotor unmanned aerial vehicle with scheduled performance. Eng. Appl. Artif. Intell. 131, 107832 (2024) https:\/\/doi.org\/10.1016\/j.engappai.2023.107832. Accessed 2025-03-26","DOI":"10.1016\/j.engappai.2023.107832"},{"key":"2309_CR38","doi-asserted-by":"publisher","unstructured":"Hickling, T., Aouf, N., Spencer, P.: Robust Adversarial Attacks Detection Based on Explainable Deep Reinforcement Learning for UAV Guidance and Planning. IEEE Trans. Intell. Veh. 8(10), 4381\u20134394 (2023). https:\/\/doi.org\/10.1109\/TIV.2023.3296227. Accessed 2024-08-27","DOI":"10.1109\/TIV.2023.3296227"},{"key":"2309_CR39","doi-asserted-by":"publisher","unstructured":"Hickling, T., Aouf, N.: Real-Time Detection of White and Black Box Adversarial Attacks on Deep Reinforcement Learning UAV Guidance. In: 2024 IEEE International Conference on Robotics and Biomimetics (ROBIO), pp. 420\u2013425 (2024). https:\/\/doi.org\/10.1109\/ROBIO64047.2024.10907751. ISSN: 2994-3574. https:\/\/ieeexplore.ieee.org\/document\/10907751\/ Accessed 2025-06-30","DOI":"10.1109\/ROBIO64047.2024.10907751"},{"key":"2309_CR40","doi-asserted-by":"publisher","unstructured":"Clark, G., Doran, M., Glisson, W.: A Malicious Attack on the Machine Learning Policy of a Robotic System. In: 2018 17th IEEE International Conference On Trust, Security And Privacy In Computing And Communications\/ 12th IEEE International Conference On Big Data Science And Engineering (TrustCom\/BigDataSE), pp. 516\u2013521 (2018). https:\/\/doi.org\/10.1109\/TrustCom\/BigDataSE.2018.00079. ISSN: 2324-9013. https:\/\/ieeexplore.ieee.org\/document\/8455947 Accessed 2025-06-30","DOI":"10.1109\/TrustCom\/BigDataSE.2018.00079"},{"key":"2309_CR41","doi-asserted-by":"publisher","unstructured":"Dasgupta, S., Ghosh, T., Rahman, M.: A Reinforcement Learning Approach for GNSS Spoofing Attack Detection of Autonomous Vehicles. Transport. Res. Record: J. Transport. Res. Board 2676(12), 318\u2013330 (2022) https:\/\/doi.org\/10.1177\/03611981221095509. arXiv:2108.08628 [eess]. Accessed 2025-06-30","DOI":"10.1177\/03611981221095509"},{"key":"2309_CR42","doi-asserted-by":"publisher","unstructured":"Wang, Z., Aouf, N.: Efficient adversarial attacks detection for deep reinforcement learning-based autonomous planetary landing GNC. Acta Astronaut. 224, 37\u201347 (2024) https:\/\/doi.org\/10.1016\/j.actaastro.2024.07.052. Accessed 2025-06-30","DOI":"10.1016\/j.actaastro.2024.07.052"},{"key":"2309_CR43","doi-asserted-by":"publisher","unstructured":"Hosseini, A., Houti, S., Qadir, J.: Deep Reinforcement Learning for Autonomous Navigation on Duckietown Platform: Evaluation of Adversarial Robustness. In: 2023 International Symposium on Networks, Computers and Communications (ISNCC), pp. 1\u20136 (2023). https:\/\/doi.org\/10.1109\/ISNCC58260.2023.10323905. ISSN: 2768-0940. https:\/\/ieeexplore.ieee.org\/document\/10323905 Accessed 2025-06-30","DOI":"10.1109\/ISNCC58260.2023.10323905"},{"key":"2309_CR44","doi-asserted-by":"publisher","unstructured":"Zhang, Z., Duan, T., Lin, Z., Huang, D., Fang, Z., Sun, Z., Xiong, L., Liang, H., Cui, H., Cui, Y., Gao, Y.: Robust Deep Reinforcement Learning in Robotics via Adaptive Gradient-Masked Adversarial Attacks. arXiv:2503.20844 [cs] (2025). https:\/\/doi.org\/10.48550\/arXiv.2503.20844. Accessed 2025-06-30","DOI":"10.48550\/arXiv.2503.20844"},{"key":"2309_CR45","unstructured":"L\u00fctjens, B., Everett, M., How, J.P.: Certified Adversarial Robustness for Deep Reinforcement Learning. arXiv:1910.12908 [cs] (2020). 10.48550\/arXiv.1910.12908 . Accessed 2025-06-30"},{"key":"2309_CR46","doi-asserted-by":"publisher","unstructured":"Ibrahum, A.D.M., Hussain, M., Hong, J.-E.: Deep learning adversarial attacks and defenses in autonomous vehicles: a systematic literature review from a safety perspective. Artif. Intell. Rev. 58(1), 28 (2024). https:\/\/doi.org\/10.1007\/s10462-024-11014-8. Accessed 2025-06-30","DOI":"10.1007\/s10462-024-11014-8"},{"key":"2309_CR47","doi-asserted-by":"publisher","unstructured":"Deng, Y., Zhang, T., Lou, G., Zheng, X., Jin, J., Han, Q.-L.: Deep Learning-Based Autonomous Driving Systems: A Survey of Attacks and Defenses. IEEE Trans. Industr. Inf. 17(12), 7897\u20137912 (2021). https:\/\/doi.org\/10.1109\/TII.2021.3071405. Accessed 2025-06-30","DOI":"10.1109\/TII.2021.3071405."},{"key":"2309_CR48","doi-asserted-by":"publisher","unstructured":"Ilahi, I., Usama, M., Qadir, J., Janjua, M.U., Al-Fuqaha, A., Hoang, D.T., Niyato, D.: Challenges and Countermeasures for Adversarial Attacks on Deep Reinforcement Learning. arXiv. arXiv:2001.09684 [cs] (2021). https:\/\/doi.org\/10.48550\/arXiv.2001.09684. Accessed 2025-06-30","DOI":"10.48550\/arXiv.2001.09684"},{"key":"2309_CR49","doi-asserted-by":"publisher","unstructured":"Standen, M., Kim, J., Szabo, C.: Adversarial Machine Learning Attacks and Defences in Multi-Agent Reinforcement Learning. ACM Comput. Surv. 57(5), 1\u201335 (2025). https:\/\/doi.org\/10.1145\/3708320. Accessed 2025-05-28","DOI":"10.1145\/3708320"},{"key":"2309_CR50","unstructured":"Korkmaz, E.: Deep Reinforcement Learning Policies Learn Shared Adversarial Features Across MDPs. arXiv. arXiv:2112.09025 [cs, stat] (2021). Accessed 2024-09-28"},{"key":"2309_CR51","unstructured":"Korkmaz, E., Brown-Cohen, J.: Detecting Adversarial Directions in Deep Reinforcement Learning to Make Robust Decisions. arXiv:2306.05873 [cs, stat] (2023). Accessed 2024-09-29"},{"key":"2309_CR52","unstructured":"Towers, M., Kwiatkowski, A., Terry, J., Balis, J.U., De\u00a0Cola, G., Deleu, T., Goul\u00e3o, M., Kallinteris, A., Krimmel, M., KG, A., Perez-Vicente, R., Pierr\u00e9, A., Schulhoff, S., Tai, J.J., Tan, H., Younis, O.G.: Gymnasium: A Standard Interface for Reinforcement Learning Environments. arXiv. arXiv:2407.17032 [cs] (2024). Accessed 2024-08-27"},{"key":"2309_CR53","doi-asserted-by":"publisher","unstructured":"Kim, S.-G., Lee, E., Hong, I.-P., Yook, J.-G.: Review of Intentional Electromagnetic Interference on UAV Sensor Modules and Experimental Study. Sensors 22(6), 2384 (2022). https:\/\/doi.org\/10.3390\/s22062384. Number: 6 Publisher: Multidisciplinary Digital Publishing Institute. Accessed 2023-07-27","DOI":"10.3390\/s22062384"},{"key":"2309_CR54","unstructured":"Fujimoto, S., Hoof, H., Meger, D.: Addressing Function Approximation Error in Actor-Critic Methods. arXiv:1802.09477 [cs, stat] (2018). Accessed 2024-02-18"},{"key":"2309_CR55","unstructured":"Schulman, J., Wolski, F., Dhariwal, P., Radford, A., Klimov, O.: Proximal Policy Optimization Algorithms. arXiv. arXiv:1707.06347 [cs] (2017). Accessed 2022-11-03"},{"key":"2309_CR56","doi-asserted-by":"publisher","unstructured":"Hochreiter, S., Schmidhuber, J.: Long Short-Term Memory. Neural Comput. 9(8), 1735\u20131780 (1997). https:\/\/doi.org\/10.1162\/neco.1997.9.8.1735. Accessed 2024-03-28","DOI":"10.1162\/neco.1997.9.8.1735"},{"key":"2309_CR57","unstructured":"Hafner, D., Pasukonis, J., Ba, J., Lillicrap, T.: Mastering Diverse Domains through World Models. arXiv. arXiv:2301.04104 [cs, stat] (2023). Accessed 2023-12-06"},{"key":"2309_CR58","unstructured":"Raffin, A., Hill, A., Gleave, A., Kanervisto, A., Ernestus, M., Dormann, N.: Stable-Baselines3: Reliable Reinforcement Learning Implementations. J. Mach. Learn. Res. 22(268), 1\u20138 (2021). Accessed 2024-10-09"},{"key":"2309_CR59","doi-asserted-by":"publisher","unstructured":"Wisniewski, M., Rana, Z.A., Petrunin, I., Holt, A., Harman, S.: Towards Fully Autonomous Drone Tracking by a Reinforcement Learning Agent Controlling a Pan\u2013Tilt\u2013Zoom Camera. Drones 8(6), 235 (2024). https:\/\/doi.org\/10.3390\/drones8060235. Number: 6 Publisher: Multidisciplinary Digital Publishing Institute. Accessed 2024-08-23","DOI":"10.3390\/drones8060235"},{"key":"2309_CR60","doi-asserted-by":"publisher","unstructured":"Zahavy, T., Zrihem, N.B., Mannor, S.: Graying the black box: Understanding DQNs. arXiv. arXiv:1602.02658 [cs] (2017). https:\/\/doi.org\/10.48550\/arXiv.1602.02658. Accessed 2025-02-13","DOI":"10.48550\/arXiv.1602.02658"},{"key":"2309_CR61","unstructured":"DeMay, C.R., White, E.L., Dunham, W.D., Pino, J.A.: AlphaDogfight Trials: Bringing Autonomy to Air Combat. Johns Hopkins APL Technical Digest 36(2) (2022)"},{"key":"2309_CR62","unstructured":"Hafner, D., Lee, K.-H., Fischer, I., Abbeel, P.: Deep Hierarchical Planning from Pixels. arXiv. arXiv:2206.04114 [cs, stat] (2022). Accessed 2024-04-24"},{"key":"2309_CR63","doi-asserted-by":"publisher","unstructured":"Pope, A.P., Ide, J.S., Mi\u0107ovi\u0107, D., Diaz, H., Twedt, J.C., Alcedo, K., Walker, T.T., Rosenbluth, D., Ritholtz, L., Javorsek, D.: Hierarchical Reinforcement Learning for Air Combat at DARPA\u2019s AlphaDogfight Trials. IEEE Trans. Artif. Intell. 4(6), 1371\u20131385 (2023). https:\/\/doi.org\/10.1109\/TAI.2022.3222143. Conference Name: IEEE Transactions on Artificial Intelligence. Accessed 2025-04-03","DOI":"10.1109\/TAI.2022.3222143"},{"key":"2309_CR64","doi-asserted-by":"publisher","unstructured":"Lechner, M., Hasani, R., Amini, A., Henzinger, T.A., Rus, D., Grosu, R.: Neural circuit policies enabling auditable autonomy. Nature Mach. Intell. 2(10), 642\u2013652 (2020). https:\/\/doi.org\/10.1038\/s42256-020-00237-3. Number: 10 Publisher: Nature Publishing Group. Accessed 2024-01-31","DOI":"10.1038\/s42256-020-00237-3"},{"key":"2309_CR65","unstructured":"Hasani, R., Lechner, M., Amini, A., Rus, D., Grosu, R.: Liquid Time-constant Networks. arXiv. arXiv:2006.04439 [cs, stat] (2020). Accessed 2024-01-17"},{"key":"2309_CR66","doi-asserted-by":"publisher","unstructured":"Chahine, M., Hasani, R., Kao, P., Ray, A., Shubert, R., Lechner, M., Amini, A., Rus, D.: Robust flight navigation out of distribution with liquid neural networks. Sci. Robot. 8(77), 8892 (2023). https:\/\/doi.org\/10.1126\/scirobotics.adc8892. Publisher: American Association for the Advancement of Science. Accessed 2025-03-26","DOI":"10.1126\/scirobotics.adc8892"},{"key":"2309_CR67","doi-asserted-by":"crossref","unstructured":"Tremblay, J., Prakash, A., Acuna, D., Brophy, M., Jampani, V., Anil, C., To, T., Cameracci, E., Boochoon, S., Birchfield, S.: Training Deep Networks with Synthetic Data: Bridging the Reality Gap by Domain Randomization. arXiv:1804.06516 [cs] (2018). Accessed 2021-07-13","DOI":"10.1109\/CVPRW.2018.00143"},{"key":"2309_CR68","doi-asserted-by":"publisher","unstructured":"Wisniewski, M., Rana, Z.A., Petrunin, I., Holt, A., Harman, S.: Drone Detection using Deep Neural Networks Trained on Pure Synthetic Data. arXiv. arXiv:2411.09077 [cs] (2024). https:\/\/doi.org\/10.48550\/arXiv.2411.09077. Accessed 2025-07-02","DOI":"10.48550\/arXiv.2411.09077"}],"container-title":["Journal of Intelligent &amp; Robotic Systems"],"original-title":[],"language":"en","link":[{"URL":"https:\/\/link.springer.com\/content\/pdf\/10.1007\/s10846-025-02309-1.pdf","content-type":"application\/pdf","content-version":"vor","intended-application":"text-mining"},{"URL":"https:\/\/link.springer.com\/article\/10.1007\/s10846-025-02309-1\/fulltext.html","content-type":"text\/html","content-version":"vor","intended-application":"text-mining"},{"URL":"https:\/\/link.springer.com\/content\/pdf\/10.1007\/s10846-025-02309-1.pdf","content-type":"application\/pdf","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2025,10,6]],"date-time":"2025-10-06T04:13:51Z","timestamp":1759724031000},"score":1,"resource":{"primary":{"URL":"https:\/\/link.springer.com\/10.1007\/s10846-025-02309-1"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2025,9,15]]},"references-count":68,"journal-issue":{"issue":"3","published-online":{"date-parts":[[2025,9]]}},"alternative-id":["2309"],"URL":"https:\/\/doi.org\/10.1007\/s10846-025-02309-1","relation":{},"ISSN":["1573-0409"],"issn-type":[{"value":"1573-0409","type":"electronic"}],"subject":[],"published":{"date-parts":[[2025,9,15]]},"assertion":[{"value":"21 October 2024","order":1,"name":"received","label":"Received","group":{"name":"ArticleHistory","label":"Article History"}},{"value":"2 September 2025","order":2,"name":"accepted","label":"Accepted","group":{"name":"ArticleHistory","label":"Article History"}},{"value":"15 September 2025","order":3,"name":"first_online","label":"First Online","group":{"name":"ArticleHistory","label":"Article History"}}],"article-number":"103"}}