{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2026,4,13]],"date-time":"2026-04-13T21:44:29Z","timestamp":1776116669961,"version":"3.50.1"},"reference-count":49,"publisher":"MDPI AG","issue":"22","license":[{"start":{"date-parts":[[2022,11,10]],"date-time":"2022-11-10T00:00:00Z","timestamp":1668038400000},"content-version":"vor","delay-in-days":0,"URL":"https:\/\/creativecommons.org\/licenses\/by\/4.0\/"}],"funder":[{"DOI":"10.13039\/501100001809","name":"National Natural Science Foundation of China","doi-asserted-by":"publisher","award":["No. 62271500"],"award-info":[{"award-number":["No. 62271500"]}],"id":[{"id":"10.13039\/501100001809","id-type":"DOI","asserted-by":"publisher"}]},{"DOI":"10.13039\/501100001809","name":"National Natural Science Foundation of China","doi-asserted-by":"publisher","award":["2020JM-343"],"award-info":[{"award-number":["2020JM-343"]}],"id":[{"id":"10.13039\/501100001809","id-type":"DOI","asserted-by":"publisher"}]},{"name":"Natural Science Foundation of Shaanxi Province","award":["No. 62271500"],"award-info":[{"award-number":["No. 62271500"]}]},{"name":"Natural Science Foundation of Shaanxi Province","award":["2020JM-343"],"award-info":[{"award-number":["2020JM-343"]}]}],"content-domain":{"domain":[],"crossmark-restriction":false},"short-container-title":["Sensors"],"abstract":"<jats:p>Airborne radars are susceptible to a large number of clutter, noise and variable jamming signals in the real environment, especially when faced with active main lobe jamming, as the waveform shortcut technology in the traditional regime can no longer meet the actual battlefield radar anti-jamming requirements. Therefore, it is necessary to study anti-main-lobe jamming techniques for airborne radars in complex environments to improve their battlefield survivability. In this paper, we propose an airborne radar waveform design method based on a deep reinforcement learning (DRL) algorithm under clutter and jamming conditions, after previous research on reinforcement-learning (RL)-based airborne radar anti-jamming waveform design methods that have improved the anti-jamming performance of airborne radars. The method uses a Markov decision process (MDP) to describe the complex operating environment of airborne radars, calculates the value of the radar anti-jamming waveform strategy under various jamming states using deep neural networks and designs the optimal anti-jamming waveform strategy for airborne radars based on the duelling double deep Q network (D3QN) algorithm. In addition, the method uses an iterative transformation method (ITM) to generate the time domain signals of the optimal waveform strategy. Simulation results show that the airborne radar waveform designed based on the deep reinforcement learning algorithm proposed in this paper improves the signal-to-jamming plus noise ratio (SJNR) by 2.08 dB and 3.03 dB, and target detection probability by 26.79% and 44.25%, respectively, compared with the waveform designed based on the reinforcement learning algorithm and the conventional linear frequency modulation (LFM) signal at a radar transmit power of 5 W. The airborne radar waveform design method proposed in this paper helps airborne radars to enhance anti-jamming performance in complex environments while further improving target detection performance.<\/jats:p>","DOI":"10.3390\/s22228689","type":"journal-article","created":{"date-parts":[[2022,11,10]],"date-time":"2022-11-10T21:22:02Z","timestamp":1668115322000},"page":"8689","update-policy":"https:\/\/doi.org\/10.3390\/mdpi_crossmark_policy","source":"Crossref","is-referenced-by-count":18,"title":["Airborne Radar Anti-Jamming Waveform Design Based on Deep Reinforcement Learning"],"prefix":"10.3390","volume":"22","author":[{"given":"Zexin","family":"Zheng","sequence":"first","affiliation":[{"name":"Information and Navigation College, Air Force Engineering University, Xi\u2019an 710077, China"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Wei","family":"Li","sequence":"additional","affiliation":[{"name":"Information and Navigation College, Air Force Engineering University, Xi\u2019an 710077, China"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Kun","family":"Zou","sequence":"additional","affiliation":[{"name":"Information and Navigation College, Air Force Engineering University, Xi\u2019an 710077, China"}],"role":[{"role":"author","vocabulary":"crossref"}]}],"member":"1968","published-online":{"date-parts":[[2022,11,10]]},"reference":[{"key":"ref_1","unstructured":"Li, Z., Tang, B., Zhou, Q., Shi, J., and Zhang, J. (2022). A review of research on the optimal design of new system airborne radar waveforms. Syst. Eng. Electron. Technol., 1\u201320. Available online: http:\/\/kns.cnki.net\/kcms\/detail\/11.2422.TN.20220328.2210.037.html."},{"key":"ref_2","unstructured":"Zhao, G. (2015). Principles of Radar Countermeasures, Xidian University Press. [2nd ed.]."},{"key":"ref_3","unstructured":"Skolnik, M. (2010). Translated by Nanjing Electronics Technology Research Institute. Radar Handbook, Electronic Industry Press."},{"key":"ref_4","unstructured":"Chen, X., Xue, Y., Zhang, L., and Huang, Y. (2021). Airborne Radar System and Information Processing, Electronic Industry Press."},{"key":"ref_5","doi-asserted-by":"crossref","unstructured":"Haykin, S. (2005, January 13\u201315). Cognitive radar networks. Proceedings of the 2005 1st IEEE International Workshop on Computational Advances in Multi-Sensor Adaptive Processing, Puerta Vallarta, Mexico.","DOI":"10.1109\/CAMAP.2005.1574168"},{"key":"ref_6","doi-asserted-by":"crossref","unstructured":"Yan, Y., Chen, H., and Su, J. (2021, January 19\u201321). Overview of anti-jamming technology in the main lobe of radar. Proceedings of the 2021 IEEE 4th International Conference on Automation, Electronics and Electrical Engineering (AUTEEE), Shenyang, China.","DOI":"10.1109\/AUTEEE52864.2021.9668666"},{"key":"ref_7","first-page":"1506","article-title":"Dense False Targets Jamming Suppression Algorithm Based on the Frequency Agility and Waveform Entropy","volume":"43","author":"Fang","year":"2021","journal-title":"Syst. Eng. Electron."},{"key":"ref_8","first-page":"3126","article-title":"Status and Prospects of Frequency Agile Radar Waveform Countermeasure Technology","volume":"43","author":"Quan","year":"2021","journal-title":"Syst. Eng. Electron."},{"key":"ref_9","first-page":"31","article-title":"Research and Test of Radar Low Intercept Waveform Anti-Main Lobe Jamming Technology","volume":"50","author":"Yan","year":"2021","journal-title":"Fire Control. Radar Technol."},{"key":"ref_10","first-page":"1","article-title":"Joint frequency and PRF agility waveform optimization for high-resolution ISAR imaging","volume":"60","author":"Wei","year":"2022","journal-title":"IEEE Trans. Geosci. Remote Sens."},{"key":"ref_11","doi-asserted-by":"crossref","unstructured":"Ou, J., Li, J., Zhang, J., and Zhan, R. (2020). Frequency Agility Radar Signal Processing, Science Press.","DOI":"10.1155\/2020\/4673763"},{"key":"ref_12","doi-asserted-by":"crossref","unstructured":"Thornton, C., Buehrer, R.M., and Martone, A. (2021). Constrained contextual andit learning for adaptive radar waveform selection. IEEE Trans. Aerosp. Electron. Syst.","DOI":"10.1109\/TAES.2021.3109110"},{"key":"ref_13","unstructured":"Cui, G., De, M., Farina, A., and Li, J. (2020). Radar Waveform Design Based on Optimization Theory, The Institution of Engineering and Technology."},{"key":"ref_14","first-page":"1210","article-title":"Anti-simultaneous jamming based on phase-coded waveform shortcuts and CFAR techniques","volume":"44","author":"Xia","year":"2022","journal-title":"Syst. Eng. Electron."},{"key":"ref_15","unstructured":"Hu, H., Lui, R., and Zhang, J. (2018, January 13\u201315). An improved radar anti-jamming method. Proceedings of the 2018 IEEE 3rd International Conference on Signal and Image Processing, Shenzhen, China."},{"key":"ref_16","doi-asserted-by":"crossref","unstructured":"Wang, H., Yang, X., and Li, Y. (2019, January 11\u201313). Radar anti-retransmitted jamming technology based on agility waveforms. Proceedings of the 2019 IEEE International Conference on Signal, Information and Data Processing (ICSIDP), Chongqing, China.","DOI":"10.1109\/ICSIDP47821.2019.9173333"},{"key":"ref_17","doi-asserted-by":"crossref","unstructured":"Zhou, C., Liu, F., and Liu, Q. (2017). An Adaptive Transmitting Scheme for Interrupted Sampling Repeater Jamming Suppression. Sensors, 17.","DOI":"10.3390\/s17112480"},{"key":"ref_18","doi-asserted-by":"crossref","first-page":"7162","DOI":"10.3390\/s110707162","article-title":"Phase-Modulated Waveform Design for Extended Target Detection in the Presence of Clutter","volume":"11","author":"Gong","year":"2011","journal-title":"Sensors"},{"key":"ref_19","first-page":"2216","article-title":"A cognitive-based approach to the design of anti-folding extended clutter waveforms","volume":"40","author":"Zhang","year":"2018","journal-title":"Syst. Eng. Electron."},{"key":"ref_20","first-page":"2448","article-title":"Cognitive constant-mode waveform design method for resisting intermittent sampling and forwarding jamming","volume":"43","author":"He","year":"2021","journal-title":"Syst. Eng. Electron."},{"key":"ref_21","first-page":"70","article-title":"Detection-oriented low peak-to-average ratio robust waveform design for cognitive radar","volume":"37","author":"Hao","year":"2022","journal-title":"Electron. Inf. Warf. Technol."},{"key":"ref_22","first-page":"726","article-title":"A Review of Cognitive Waveform Optimisation for Different Radar Tasks","volume":"50","author":"Yu","year":"2022","journal-title":"Acta Electron. Sin."},{"key":"ref_23","doi-asserted-by":"crossref","first-page":"10181","DOI":"10.3390\/s101110181","article-title":"Extended Target Recognition in Cognitive Radar Networks","volume":"10","author":"Wei","year":"2010","journal-title":"Sensors"},{"key":"ref_24","unstructured":"Wei, Y., and Yuan, Y. (2020, January 21\u201325). Reinforcement learning-based joint adaptive frequency hopping and pulse-width allocation for radar anti-jamming. Proceedings of the 2020 IEEE Radar Conference, Florence, Italy."},{"key":"ref_25","doi-asserted-by":"crossref","unstructured":"Liu, K., Lu, X., Xiao, L., and Xu, L. (2020, January 7\u201311). Learning based energy efficient radar power control against deceptive jamming. Proceedings of the 2020 IEEE Global Communications Conference, Taipei, Taiwan.","DOI":"10.1109\/GLOBECOM42002.2020.9322520"},{"key":"ref_26","unstructured":"Wang, H., and Wang, F. (2020). Application of a reinforcement learning algorithm in intelligent anti-jamming of radar. Mod. Radar, 42."},{"key":"ref_27","doi-asserted-by":"crossref","first-page":"2622","DOI":"10.1109\/TAES.2021.3061809","article-title":"A reinforcement learning based approach for multitarget detection in massive MIMO radar","volume":"57","author":"Ahmed","year":"2021","journal-title":"IEEE Trans. Aerosp. Electron. Syst."},{"key":"ref_28","unstructured":"Li, K., Jiu, B., Liu, H., and Liang, S. (2018, January 14\u201316). Reinforcement Learning based Anti-jamming Frequency Hopping Strategies Design for Cognitive Radar. Proceedings of the 2018 IEEE International Conference on Signal Processing, Communications and Computing (ICSPCC), Qingdao, China."},{"key":"ref_29","doi-asserted-by":"crossref","unstructured":"Li, K., Jiu, B., and Liu, H. (October, January 30). Deep Q-network based anti-jamming strategy design for frequency agile radar. Proceedings of the 2019 International Radar Conference, Toulon, France.","DOI":"10.1109\/RADAR41533.2019.171227"},{"key":"ref_30","doi-asserted-by":"crossref","first-page":"108130","DOI":"10.1016\/j.sigpro.2021.108130","article-title":"Radar Active Antagonism through Deep Reinforcement Learning: A Way to Address The Challenge of Mainlobe Jamming","volume":"186","author":"Li","year":"2021","journal-title":"Signal Process."},{"key":"ref_31","doi-asserted-by":"crossref","unstructured":"Ak, S., and Br\u00fcggenwirth, S. (2020, January 28\u201330). Avoiding Jammers: A Reinforcement Learning Approach. Proceedings of the 2020 IEEE International Radar Conferenc (RADAR), Washington, DC, USA.","DOI":"10.1109\/RADAR42522.2020.9114797"},{"key":"ref_32","doi-asserted-by":"crossref","unstructured":"Selvi, E., Buehrer, R., Martone, A., and Sherbondy, K. (2018, January 23\u201327). On The Use of Markov Decision Processes in Cognitive Radar: An Application to Target Tracking. Proceedings of the 2018 IEEE Radar Conference (RadarConf18), Oklahoma City, OK, USA.","DOI":"10.1109\/RADAR.2018.8378616"},{"key":"ref_33","doi-asserted-by":"crossref","unstructured":"Kozy, M., Yu, J., Buehrer, R., Martone, A., and Sherbondy, K. (2019, January 22\u201326). Applying Deep-Q Networks to Target Tracking to Improve Cognitive Radar. Proceedings of the 2019 IEEE Radar Conference (RadarConf), Boston, MA, USA.","DOI":"10.1109\/RADAR.2019.8835780"},{"key":"ref_34","doi-asserted-by":"crossref","first-page":"1264","DOI":"10.1109\/JSYST.2020.2984774","article-title":"Data-Driven Simultaneous Multibeam Power Allocation: When Multiple Targets Tracking Meets Deep Reinforcement Learning","volume":"15","author":"Shi","year":"2020","journal-title":"IEEE Syst. J."},{"key":"ref_35","doi-asserted-by":"crossref","unstructured":"Hessel, M., Modayil, J., Hasselt, H., Schaul, T., Ostrovski, G., Dabney, W., Horgan, D., Piot, B., Azar, M., and Silver, D. (2018, January 2\u20137). Rainbow: Combining improvements in deep reinforcement learning. Proceedings of the 32rd AAAI Conference on Artificial Intelligence, New Orleans, LA, USA.","DOI":"10.1609\/aaai.v32i1.11796"},{"key":"ref_36","doi-asserted-by":"crossref","first-page":"279","DOI":"10.1007\/BF00992698","article-title":"Q-learning","volume":"8","author":"Watkins","year":"1992","journal-title":"Mach. Learn"},{"key":"ref_37","unstructured":"Mnih, V., Kavukcuoglu, K., Silver, D., Graves, A., Antonoglou, I., Wierstra, D., and Riedmiller, M. (2013). Playing Atari with Deep Reinforcement Learning. arXiv, Available online: https:\/\/arxiv.org\/pdf\/1312.5602.pdf."},{"key":"ref_38","unstructured":"Bellemare, M.G., Dabney, W., and Munos, R. (2017, January 6\u201311). A Distributional Perspective on Reinforcement Learning. Proceedings of the 34rd International Conference on Machine Learning (ICML 2017), Sydney, Australia. Available online: https:\/\/arxiv.org\/pdf\/1707.06887.pdf."},{"key":"ref_39","doi-asserted-by":"crossref","unstructured":"Hasselt, V.H., Guez, A., and Silver, D. (2016, January 12\u201317). Deep Reinforcement Learning with Double Q-Learning. The 30rd AAAI Conference on Artificial Intelligence, Phoenix, AZ, USA. Available online: https:\/\/arxiv.org\/pdf\/1509.06461v3.pdf.","DOI":"10.1609\/aaai.v30i1.10295"},{"key":"ref_40","unstructured":"Schaul, T., Quan, J., Antonoglou, I., and Silver, D. (2015). Prioritized Experience Replay. arXiv, Available online: https:\/\/arxiv.org\/pdf\/1511.05952.pdf."},{"key":"ref_41","unstructured":"Wang, Z., Schaul, T., Hessel, M., Hasselt, H., Lanctot, M., and Freitas, N. (2016, January 19\u201324). Dueling Network Architectures for Deep Reinforcement Learning. Proceedings of the 33rd International Conference on Machine Learning (ICML 2016), New York, NY, USA. Available online: https:\/\/arxiv.org\/pdf\/1511.06581.pdf."},{"key":"ref_42","unstructured":"Fortunato, M., Azar, M.G., Piot, B., Menick, J., Osband, I., Graves, A., Mnih, V., Munos, R., Hassabis, D., and Pietquin, O. (2017). Noisy Networks for Exploration. arXiv, Available online: https:\/\/arxiv.org\/pdf\/1706.10295.pdf."},{"key":"ref_43","unstructured":"Zheng, Z., Li, W., Zou, K., and Li, Y. (2022). Anti-jamming Waveform Design of Ground-to-air Radar Based on Reinforcement Learning. Acta Armamentarii, 1\u20139. Available online: http:\/\/kns.cnki.net\/kcms\/detail\/11.2176.TJ.20220722.1003.002.html."},{"key":"ref_44","first-page":"757","article-title":"Radar Waveform Strategy Based on Game Theory","volume":"28","author":"Wang","year":"2019","journal-title":"Radio Eng."},{"key":"ref_45","doi-asserted-by":"crossref","first-page":"912","DOI":"10.1109\/TAES.2011.5751234","article-title":"Theory and Application of SNR and Mutual Information Matched Illumination Waveforms","volume":"47","author":"Romero","year":"2011","journal-title":"IEEE Trans. Aerosp. Electron. Syst."},{"key":"ref_46","doi-asserted-by":"crossref","unstructured":"Gagniuc, P. (2017). Markov Chains: From Theory to Implementation and Experimentation, Wiley.","DOI":"10.1002\/9781119387596"},{"key":"ref_47","unstructured":"Steven, M.K. (2014). Fundamentals of Statistical Signal Processing: Estimation and Detection Theory, Electronic Industry Press."},{"key":"ref_48","first-page":"1863","article-title":"Research Advance on Cognitive Radar and Its Key Technology","volume":"40","author":"Li","year":"2012","journal-title":"Acta Electron. Sin."},{"key":"ref_49","doi-asserted-by":"crossref","first-page":"910","DOI":"10.1109\/TAES.2010.5461666","article-title":"Iterative Method for Nonlinear FM Synthesis of Radar Signals","volume":"46","author":"Jackson","year":"2010","journal-title":"IEEE Trans. Aerosp. Electron. Syst."}],"container-title":["Sensors"],"original-title":[],"language":"en","link":[{"URL":"https:\/\/www.mdpi.com\/1424-8220\/22\/22\/8689\/pdf","content-type":"unspecified","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2025,10,11]],"date-time":"2025-10-11T01:13:54Z","timestamp":1760145234000},"score":1,"resource":{"primary":{"URL":"https:\/\/www.mdpi.com\/1424-8220\/22\/22\/8689"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2022,11,10]]},"references-count":49,"journal-issue":{"issue":"22","published-online":{"date-parts":[[2022,11]]}},"alternative-id":["s22228689"],"URL":"https:\/\/doi.org\/10.3390\/s22228689","relation":{},"ISSN":["1424-8220"],"issn-type":[{"value":"1424-8220","type":"electronic"}],"subject":[],"published":{"date-parts":[[2022,11,10]]}}}