{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2026,5,19]],"date-time":"2026-05-19T04:20:13Z","timestamp":1779164413757,"version":"3.51.4"},"reference-count":37,"publisher":"Springer Science and Business Media LLC","issue":"1","license":[{"start":{"date-parts":[[2025,12,29]],"date-time":"2025-12-29T00:00:00Z","timestamp":1766966400000},"content-version":"tdm","delay-in-days":0,"URL":"https:\/\/creativecommons.org\/licenses\/by\/4.0"},{"start":{"date-parts":[[2025,12,29]],"date-time":"2025-12-29T00:00:00Z","timestamp":1766966400000},"content-version":"vor","delay-in-days":0,"URL":"https:\/\/creativecommons.org\/licenses\/by\/4.0"}],"funder":[{"DOI":"10.13039\/501100001809","name":"National Natural Science Foundation of China","doi-asserted-by":"publisher","award":["62472010"],"award-info":[{"award-number":["62472010"]}],"id":[{"id":"10.13039\/501100001809","id-type":"DOI","asserted-by":"publisher"}]},{"name":"Chongging Natural Science Foundation","award":["CSTB2022NSCQ-MSX1415"],"award-info":[{"award-number":["CSTB2022NSCQ-MSX1415"]}]},{"name":"Ministry of Education Foundation on Humanities and Social Sciences","award":["23YJAZH129"],"award-info":[{"award-number":["23YJAZH129"]}]},{"name":"Chongging Natural Science Foundation","award":["CSTB2024NSCQ-MSX0687"],"award-info":[{"award-number":["CSTB2024NSCQ-MSX0687"]}]}],"content-domain":{"domain":["link.springer.com"],"crossmark-restriction":false},"short-container-title":["Complex Intell. Syst."],"published-print":{"date-parts":[[2026,1]]},"abstract":"<jats:title>Abstract<\/jats:title>\n                  <jats:p>Deep reinforcement learning shows broad prospects in multi-unmanned aerial vehicle(UAV) collaborative search and rescue tasks. However, in the face of high-dimensional collaborative decision-making spaces and limited computing resources, its performance is vulnerable to limitations. This paper proposes a deep deterministic policy gradient method based on linear attention. By introducing the linear attention mechanism based on random feature mapping, while effectively modeling the interaction among UAVs, the computational and storage overcosts caused by the increase in the number of UAVs have been significantly reduced. Furthermore, by combining smooth experience replay and adaptive importance sampling mechanism, the training efficiency and strategy stability have been further improved. The simulation experiments on both post-disaster response search and dynamic containment tasks demonstrate that the proposed algorithm consistently outperforms existing methods. In small-scale scenarios, it maintains nearly perfect success rates, while in medium- and large-scale settings it achieves up to 90.6% and 85.2% success rates in the post-disaster response search task and up to 90.1% and 80.2% in the containment task, corresponding to relative improvements of 15\u201321% over baselines. These results highlight both the robustness of the method in simple cases and its clear advantage under more challenging multi-UAV conditions.<\/jats:p>","DOI":"10.1007\/s40747-025-02166-3","type":"journal-article","created":{"date-parts":[[2025,12,29]],"date-time":"2025-12-29T13:23:05Z","timestamp":1767014585000},"update-policy":"https:\/\/doi.org\/10.1007\/springer_crossmark_policy","source":"Crossref","is-referenced-by-count":1,"title":["A multi-UAV rapid post-disaster search and rescue method based on deep reinforcement learning"],"prefix":"10.1007","volume":"12","author":[{"ORCID":"https:\/\/orcid.org\/0000-0002-2786-5958","authenticated-orcid":false,"given":"Li","family":"Tan","sequence":"first","affiliation":[],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Haixia","family":"Zhao","sequence":"additional","affiliation":[],"role":[{"role":"author","vocabulary":"crossref"}]}],"member":"297","published-online":{"date-parts":[[2025,12,29]]},"reference":[{"key":"2166_CR1","doi-asserted-by":"publisher","DOI":"10.1016\/j.soildyn.2023.108370","volume":"177","author":"P Fu","year":"2024","unstructured":"Fu P, Li X, Xu L et al (2024) An advanced assessment framework for seismic resilience of railway continuous girder bridge with multiple spans considering 72h golden rescue requirements. Soil Dyn Earthq Eng 177:108370","journal-title":"Soil Dyn Earthq Eng"},{"key":"2166_CR2","doi-asserted-by":"crossref","unstructured":"Tong K, Hu Y, Dikic B, et al (2024) Robots saving lives: a literature review about search and rescue (sar) in harsh environments. In: 2024 IEEE intelligent vehicles symposium (IV). IEEE, pp 953\u2013960","DOI":"10.1109\/IV55156.2024.10588685"},{"issue":"6","key":"2166_CR3","first-page":"1799","volume":"5","author":"VG Nair","year":"2024","unstructured":"Nair VG, D\u2019Souza JM, Rafikh RM (2024) A scoping review on unmanned aerial vehicles in disaster management: challenges and opportunities. J Robotics Control 5(6):1799\u20131826","journal-title":"J Robotics Control"},{"key":"2166_CR4","doi-asserted-by":"publisher","first-page":"1","DOI":"10.31875\/2409-9694.2024.11.01","volume":"11","author":"A Petrovic","year":"2024","unstructured":"Petrovic A, Bacanin N, Jovanovic L et al (2024) Computer-vision unmanned aerial vehicle detection system using yolov8 architectures. Int J Robot Autom Technol 11:1\u201312","journal-title":"Int J Robot Autom Technol"},{"issue":"2","key":"2166_CR5","doi-asserted-by":"publisher","DOI":"10.1016\/j.cja.2024.11.011","volume":"38","author":"L Yang","year":"2025","unstructured":"Yang L, Zhang X, Li Z et al (2025) A LODBO algorithm for multi-UAV search and rescue path planning in disaster areas. Chin J Aeronaut 38(2):103301","journal-title":"Chin J Aeronaut"},{"key":"2166_CR6","doi-asserted-by":"publisher","first-page":"37","DOI":"10.31875\/2409-9694.2024.11.03","volume":"11","author":"L Jovanovic","year":"2024","unstructured":"Jovanovic L, Antonijevic M, Perisic J et al (2024) Computer vision based areal photographic rocket detection using yolov8 models. Int J Robot Autom Technol 11:37\u201349","journal-title":"Int J Robot Autom Technol"},{"issue":"5","key":"2166_CR7","doi-asserted-by":"publisher","first-page":"8848","DOI":"10.1109\/JIOT.2023.3320796","volume":"11","author":"Y Guan","year":"2023","unstructured":"Guan Y, Zou S, Peng H et al (2023) Cooperative UAV trajectory design for disaster area emergency communications: a multiagent PPO method. IEEE Internet Things J 11(5):8848\u20138859","journal-title":"IEEE Internet Things J"},{"key":"2166_CR8","doi-asserted-by":"crossref","unstructured":"Han S, Liu X, Zhou M C, et al (2024) Joint association, deployment and flight trajectory optimization for multi-UAV-enabled large-scale mobile edge computing. IEEE Trans Mobile Comput","DOI":"10.1109\/TMC.2024.3426945"},{"key":"2166_CR9","doi-asserted-by":"crossref","unstructured":"Akter S, Duong DVA, Yoon S (2025) Joint optimization of UAV trajectory, task offloading, and resource allocation in UAV-aided emergency response operations. IEEE Internet Things J","DOI":"10.1109\/JIOT.2025.3549048"},{"key":"2166_CR10","doi-asserted-by":"crossref","unstructured":"Lei H, Meng D, Ran H, et al (2024) Multi-UAV trajectory design for fair and secure communication. IEEE Trans Cognit Commun Netw","DOI":"10.1109\/TCCN.2024.3487142"},{"key":"2166_CR11","doi-asserted-by":"publisher","DOI":"10.1016\/j.asoc.2023.110592","volume":"148","author":"H Zhang","year":"2023","unstructured":"Zhang H, Ma H, Mersha BW et al (2023) Distributed cooperative search method for multi-UAV with unstable communications. Appl Soft Comput 148:110592","journal-title":"Appl Soft Comput"},{"key":"2166_CR12","doi-asserted-by":"crossref","unstructured":"Cao X, Li M, Tao Y, et al (2024) HMA-SAR: multi-agent search and rescue for unknown located dynamic targets in completely unknown environments. IEEE Robot Autom Lett","DOI":"10.1109\/LRA.2024.3396097"},{"key":"2166_CR13","doi-asserted-by":"crossref","unstructured":"Tarekegn G B, Tesfaw B A, Juang R T, et al (2025) Trajectory control and fair communications for multi-UAV networks: a federated multi-agent deep reinforcement learning approach. IEEE Trans Wirel Commun","DOI":"10.1109\/TWC.2025.3561271"},{"key":"2166_CR14","unstructured":"Lowe R, Wu Y I, Tamar A, et al (2017) Multi-agent actor-critic for mixed cooperative-competitive environments. Adv Neural Inf Process Syst 30"},{"key":"2166_CR15","unstructured":"Iqbal S, Sha F (2019) Actor-attention-critic for multi-agent reinforcement learning. In: International conference on machine learning. PMLR, pp 2961\u20132970"},{"key":"2166_CR16","unstructured":"Peng H, Pappas N, Yogatama D, et al (2021) Random feature attention. arXiv preprint arXiv:2103.02143"},{"key":"2166_CR17","doi-asserted-by":"crossref","unstructured":"Toskovic A, Petrovic A, Jovanovic L, et al (2023) Marine vessel trajectory forecasting using long short-term memory neural networks optimized via modified metaheuristic algorithm. In: International conference on trends in sustainable computing and machine intelligence. Singapore: Springer Nature Singapore, pp 51\u201366","DOI":"10.1007\/978-981-99-9436-6_5"},{"key":"2166_CR18","doi-asserted-by":"crossref","unstructured":"Protic M, Jovanovic L, Dobrojevic M, et al (2024) Signals intelligence based drone detection using YOLOv8 models. In: Proceedings of the 2nd international conference on innovation in information technology and business (ICIITB 2024). Springer Nature, 113, 74\u201386","DOI":"10.2991\/978-94-6463-482-2_6"},{"issue":"20","key":"2166_CR19","doi-asserted-by":"publisher","first-page":"17734","DOI":"10.1109\/JIOT.2023.3277850","volume":"10","author":"J Li","year":"2023","unstructured":"Li J, Xiong Y, She J (2023) UAV path planning for target coverage task in dynamic environment. IEEE Internet Things J 10(20):17734\u201317745","journal-title":"IEEE Internet Things J"},{"issue":"3","key":"2166_CR20","first-page":"583","volume":"22","author":"H Wang","year":"2021","unstructured":"Wang H, Tan L, Shi J et al (2021) An improved NSGA-II algorithm for UAV path planning problems. J Internet Technol 22(3):583\u2013592","journal-title":"J Internet Technol"},{"issue":"3","key":"2166_CR21","doi-asserted-by":"publisher","first-page":"887","DOI":"10.3390\/s24030887","volume":"24","author":"N Xing","year":"2024","unstructured":"Xing N, Zhang Y, Wang Y et al (2024) Unmanned aerial vehicle cooperative data dissemination based on graph neural networks. Sensors 24(3):887","journal-title":"Sensors"},{"issue":"3","key":"2166_CR22","doi-asserted-by":"publisher","first-page":"2123","DOI":"10.1109\/TAES.2022.3208865","volume":"59","author":"K Xiong","year":"2022","unstructured":"Xiong K, Zhang T, Cui G et al (2022) Coalition game of radar network for multitarget tracking via model-based multiagent reinforcement learning. IEEE Trans Aerosp Electron Syst 59(3):2123\u20132140","journal-title":"IEEE Trans Aerosp Electron Syst"},{"issue":"6","key":"2166_CR23","doi-asserted-by":"publisher","first-page":"5154","DOI":"10.1109\/TITS.2023.3341636","volume":"25","author":"Y He","year":"2023","unstructured":"He Y, Wang D, Huang F et al (2023) Aerial-ground integrated vehicular networks: a UAV-vehicle collaboration perspective. IEEE Trans Intell Transp Syst 25(6):5154\u20135169","journal-title":"IEEE Trans Intell Transp Syst"},{"issue":"11","key":"2166_CR24","doi-asserted-by":"publisher","first-page":"895","DOI":"10.3390\/aerospace11110895","volume":"11","author":"T Wang","year":"2024","unstructured":"Wang T, Wang Z, Li W et al (2024) Flexible combinatorial-bids-based auction for cooperative target assignment of unmanned aerial vehicles. Aerospace 11(11):895","journal-title":"Aerospace"},{"key":"2166_CR25","first-page":"4385","volume":"35","author":"Y Xiao","year":"2022","unstructured":"Xiao Y, Tan W, Amato C (2022) Asynchronous actor-critic for multi-agent reinforcement learning. Adv Neural Inf Process Syst 35:4385\u20134400","journal-title":"Adv Neural Inf Process Syst"},{"key":"2166_CR26","unstructured":"Wang Z, Wang J, Zuo D, et al (2024) A hierarchical adaptive multi-task reinforcement learning framework for multiplier circuit design. In: Forty-first international conference on machine learning"},{"issue":"3","key":"2166_CR27","doi-asserted-by":"publisher","first-page":"3185","DOI":"10.1109\/JSYST.2019.2937346","volume":"14","author":"P Yao","year":"2019","unstructured":"Yao P, Zhao Z, Zhu Q (2019) Path planning for autonomous underwater vehicles with simultaneous arrival in ocean environment. IEEE Syst J 14(3):3185\u20133193","journal-title":"IEEE Syst J"},{"key":"2166_CR28","unstructured":"Hu S, Zhu F, Chang X, et al (2021) Updet: universal multi-agent reinforcement learning via policy decoupling with transformers. arXiv preprint arXiv:2101.08001"},{"key":"2166_CR29","doi-asserted-by":"crossref","unstructured":"Feng Z, Wu D, Huang M, et al (2024) Graph attention-based reinforcement learning for trajectory design and resource assignment in multi-UAV assisted communication. IEEE Internet Things J","DOI":"10.1109\/JIOT.2024.3397823"},{"key":"2166_CR30","unstructured":"Yin H, Yang Z, Zhang L, et al (2025) Attention-augmented inverse reinforcement learning with graph convolutions for multi-agent task allocation. arXiv preprint arXiv:2504.05045"},{"key":"2166_CR31","doi-asserted-by":"crossref","unstructured":"Ryu H, Shin H, Park J (2020) Multi-agent actor-critic with hierarchical graph attention network. In: Proceedings of the AAAI conference on artificial intelligence, 34(05):7236\u20137243","DOI":"10.1609\/aaai.v34i05.6214"},{"key":"2166_CR32","first-page":"16509","volume":"35","author":"M Wen","year":"2022","unstructured":"Wen M, Kuba J, Lin R et al (2022) Multi-agent reinforcement learning is a sequence modeling problem. Adv Neural Inf Process Syst 35:16509\u201316521","journal-title":"Adv Neural Inf Process Syst"},{"issue":"10","key":"2166_CR33","doi-asserted-by":"publisher","first-page":"6851","DOI":"10.1109\/TNNLS.2022.3215774","volume":"34","author":"W Du","year":"2022","unstructured":"Du W, Ding S, Zhang C et al (2022) Multiagent reinforcement learning with heterogeneous graph attention network. IEEE Trans Neural Netw Learn Syst 34(10):6851\u20136860","journal-title":"IEEE Trans Neural Netw Learn Syst"},{"issue":"9","key":"2166_CR34","doi-asserted-by":"publisher","first-page":"11648","DOI":"10.1109\/TITS.2024.3379508","volume":"25","author":"J Wu","year":"2024","unstructured":"Wu J, Li D, Yu Y et al (2024) An attention mechanism and adaptive accuracy triple-dependent MADDPG formation control method for hybrid UAVs. IEEE Trans Intell Transp Syst 25(9):11648-11663","journal-title":"IEEE Trans Intell Transp Syst"},{"key":"2166_CR35","doi-asserted-by":"crossref","unstructured":"Sun C, Shen M, Jonathan P. How. Scaling up multiagent reinforcement learning for robotic systems: Learn an adaptive sparse communication graph. In: 2020 IEEE RSJ International Conference on Intelligent Robots and Systems (IROS), pp 11755\u201311762","DOI":"10.1109\/IROS45743.2020.9341303"},{"issue":"8","key":"2166_CR36","doi-asserted-by":"publisher","first-page":"10556","DOI":"10.1109\/TITS.2021.3094821","volume":"23","author":"J Li","year":"2021","unstructured":"Li J, Ma H, Zhang Z et al (2021) Spatio-temporal graph dual-attention network for multi-agent prediction and tracking. IEEE Trans Intell Transp Syst 23(8):10556\u201310569","journal-title":"IEEE Trans Intell Transp Syst"},{"issue":"1","key":"2166_CR37","doi-asserted-by":"publisher","first-page":"4","DOI":"10.3390\/e27010004","volume":"27","author":"T Li","year":"2024","unstructured":"Li T, Shi D, Jin S et al (2024) Multi-agent hierarchical graph attention actor\u2013critic reinforcement learning. Entropy 27(1):4","journal-title":"Entropy"}],"container-title":["Complex &amp; Intelligent Systems"],"original-title":[],"language":"en","link":[{"URL":"https:\/\/link.springer.com\/content\/pdf\/10.1007\/s40747-025-02166-3.pdf","content-type":"application\/pdf","content-version":"vor","intended-application":"text-mining"},{"URL":"https:\/\/link.springer.com\/article\/10.1007\/s40747-025-02166-3","content-type":"text\/html","content-version":"vor","intended-application":"text-mining"},{"URL":"https:\/\/link.springer.com\/content\/pdf\/10.1007\/s40747-025-02166-3.pdf","content-type":"application\/pdf","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2026,1,30]],"date-time":"2026-01-30T11:48:48Z","timestamp":1769773728000},"score":1,"resource":{"primary":{"URL":"https:\/\/link.springer.com\/10.1007\/s40747-025-02166-3"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2025,12,29]]},"references-count":37,"journal-issue":{"issue":"1","published-print":{"date-parts":[[2026,1]]}},"alternative-id":["2166"],"URL":"https:\/\/doi.org\/10.1007\/s40747-025-02166-3","relation":{},"ISSN":["2199-4536","2198-6053"],"issn-type":[{"value":"2199-4536","type":"print"},{"value":"2198-6053","type":"electronic"}],"subject":[],"published":{"date-parts":[[2025,12,29]]},"assertion":[{"value":"20 July 2025","order":1,"name":"received","label":"Received","group":{"name":"ArticleHistory","label":"Article History"}},{"value":"3 November 2025","order":2,"name":"accepted","label":"Accepted","group":{"name":"ArticleHistory","label":"Article History"}},{"value":"29 December 2025","order":3,"name":"first_online","label":"First Online","group":{"name":"ArticleHistory","label":"Article History"}}],"article-number":"41"}}