{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2026,7,7]],"date-time":"2026-07-07T15:47:19Z","timestamp":1783439239593,"version":"3.54.6"},"reference-count":40,"publisher":"Springer Science and Business Media LLC","issue":"1","license":[{"start":{"date-parts":[[2023,12,21]],"date-time":"2023-12-21T00:00:00Z","timestamp":1703116800000},"content-version":"tdm","delay-in-days":0,"URL":"https:\/\/creativecommons.org\/licenses\/by\/4.0"},{"start":{"date-parts":[[2023,12,21]],"date-time":"2023-12-21T00:00:00Z","timestamp":1703116800000},"content-version":"vor","delay-in-days":0,"URL":"https:\/\/creativecommons.org\/licenses\/by\/4.0"}],"funder":[{"DOI":"10.13039\/501100003392","name":"Natural Science Foundation of Fujian Province","doi-asserted-by":"publisher","award":["2021J011086"],"award-info":[{"award-number":["2021J011086"]}],"id":[{"id":"10.13039\/501100003392","id-type":"DOI","asserted-by":"publisher"}]},{"DOI":"10.13039\/501100003392","name":"Natural Science Foundation of Fujian Province","doi-asserted-by":"publisher","award":["2023J01964"],"award-info":[{"award-number":["2023J01964"]}],"id":[{"id":"10.13039\/501100003392","id-type":"DOI","asserted-by":"publisher"}]},{"DOI":"10.13039\/501100003392","name":"Natural Science Foundation of Fujian Province","doi-asserted-by":"publisher","award":["2023J01965"],"award-info":[{"award-number":["2023J01965"]}],"id":[{"id":"10.13039\/501100003392","id-type":"DOI","asserted-by":"publisher"}]},{"DOI":"10.13039\/501100003392","name":"Natural Science Foundation of Fujian Province","doi-asserted-by":"publisher","award":["2023J01966"],"award-info":[{"award-number":["2023J01966"]}],"id":[{"id":"10.13039\/501100003392","id-type":"DOI","asserted-by":"publisher"}]},{"DOI":"10.13039\/501100017596","name":"Natural Science Basic Research Program of Shaanxi Province","doi-asserted-by":"publisher","award":["2021JM-020"],"award-info":[{"award-number":["2021JM-020"]}],"id":[{"id":"10.13039\/501100017596","id-type":"DOI","asserted-by":"publisher"}]},{"name":"Fujian Province Chinese Academy of Sciences STS Program Supporting Project","award":["2023T3084"],"award-info":[{"award-number":["2023T3084"]}]},{"name":"Qimai Science and Technology Innovation Project of Wuping Country"}],"content-domain":{"domain":["link.springer.com"],"crossmark-restriction":false},"short-container-title":["EURASIP J. Adv. Signal Process."],"abstract":"<jats:title>Abstract<\/jats:title><jats:p>Object detection holds a crucial role in medical diagnostics. Tasks like organ segmentation and malignancy diagnosis typically necessitate preliminary localization of corresponding anatomical structures. Precise positioning ensures that only pertinent regions require processing, leading to a potential reduction in computational and storage demands. Conventional image detection approaches necessitate numerous candidate boxes, resulting in redundant computations. Developing techniques capable of accurately detecting medical image objects without reliance on candidate boxes holds substantial practical significance. This paper introduces a 2D method for detecting medical image objects, which leverages multi-agent deep Q-network reinforcement learning and a multi-scale image representation. The method constructs a collaborative environment for multiple agents. These agents individually govern the upper-right corner and lower-left corner positions of the object detection frame, progressively converging toward the actual endpoint through iterative interactions. To expedite the detection process, a multi-scale image representation technique is employed. This method segments the process into three scales. Initially, within the coarse-scale space, the agent approximates the region containing the true endpoint, subsequently executing oscillatory movements. Progressively, it refines its approach within the fine-scale space, advancing toward the genuine endpoint with smaller iterative steps. The detection results demonstrate that collaborative detection among agents yields a 2.45<jats:inline-formula><jats:alternatives><jats:tex-math>$$\\%$$<\/jats:tex-math><mml:math xmlns:mml=\"http:\/\/www.w3.org\/1998\/Math\/MathML\">\n                  <mml:mo>%<\/mml:mo>\n                <\/mml:math><\/jats:alternatives><\/jats:inline-formula> higher intersection over union compared to non-collaborative detection. Agents exhibit varying step sizes and fields of view in different scale spaces, leading to a reduction in detection time by 0.12\u00a0s compared to single-scale comparison. Experimental outcomes demonstrate the superiority of the medical image target detection method proposed in this study over prevailing mainstream detection algorithms.<\/jats:p>","DOI":"10.1186\/s13634-023-01095-y","type":"journal-article","created":{"date-parts":[[2023,12,21]],"date-time":"2023-12-21T15:03:09Z","timestamp":1703170989000},"update-policy":"https:\/\/doi.org\/10.1007\/springer_crossmark_policy","source":"Crossref","is-referenced-by-count":5,"title":["Enhancing medical image object detection with collaborative multi-agent deep Q-networks and multi-scale representation"],"prefix":"10.1186","volume":"2023","author":[{"given":"Qinghui","family":"Wang","sequence":"first","affiliation":[],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Fenglin","family":"Liu","sequence":"additional","affiliation":[],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Ruirui","family":"Zou","sequence":"additional","affiliation":[],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Ying","family":"Wang","sequence":"additional","affiliation":[],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Chenyang","family":"Zheng","sequence":"additional","affiliation":[],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Zhiqiang","family":"Tian","sequence":"additional","affiliation":[],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Shaoyi","family":"Du","sequence":"additional","affiliation":[],"role":[{"vocabulary":"crossref","role":"author"}]},{"ORCID":"https:\/\/orcid.org\/0000-0002-8353-8265","authenticated-orcid":false,"given":"Wei","family":"Zeng","sequence":"additional","affiliation":[],"role":[{"vocabulary":"crossref","role":"author"}]}],"member":"297","published-online":{"date-parts":[[2023,12,21]]},"reference":[{"key":"1095_CR1","doi-asserted-by":"crossref","unstructured":"K. He, G. Gkioxari , P. Doll\u00e1ir, R. Girshick, Mask, R-CNN, in Proceedings of the IEEE International Conference on Computer Vision (2017), pp. 2961\u20132969","DOI":"10.1109\/ICCV.2017.322"},{"key":"1095_CR2","doi-asserted-by":"crossref","unstructured":"P. Anderson, X. He, C. Buehler, D. Teney, M. Johnson, S. Gould, L. Zhang, Bottom-up and top-down attention for image captioning and visual question answering, in Proceedings of the IEEE International Conference on Computer Vision (2018), pp. 6077\u20136086","DOI":"10.1109\/CVPR.2018.00636"},{"key":"1095_CR3","doi-asserted-by":"crossref","unstructured":"J. Yang, J. Lu, S. Lee, D. Batra, D. Parikh, Graph R-CNN for scene graph generation, in Proceedings of the European Conference on Computer Vision (2018), pp. 670\u2013685","DOI":"10.1007\/978-3-030-01246-5_41"},{"issue":"8","key":"1095_CR4","doi-asserted-by":"publisher","first-page":"5791","DOI":"10.1007\/s00521-022-06960-9","volume":"34","author":"MA Abdou","year":"2022","unstructured":"M.A. Abdou, Literature review: efficient deep neural networks techniques for medical image analysis. Neural Comput. Appl. 34(8), 5791\u20135812 (2022)","journal-title":"Neural Comput. Appl."},{"key":"1095_CR5","doi-asserted-by":"crossref","unstructured":"P.N. Samarakoon, E. Promayon, C. Fouard, Light random regression forests for automatic multi-organ localization in CT images, in IEEE 14th International Symposium on Biomedical Imaging (2017), pp. 371\u2013374","DOI":"10.1109\/ISBI.2017.7950540"},{"issue":"9","key":"1095_CR6","doi-asserted-by":"publisher","first-page":"2794","DOI":"10.1109\/TMI.2020.2975853","volume":"39","author":"S Liang","year":"2020","unstructured":"S. Liang, K.H. Thung, D. Nie, Y. Zhang, D. Shen, Multi-view spatial aggregation framework for joint localization and segmentation of organs at risk in head and neck CT images. IEEE Trans. Med. Imaging 39(9), 2794\u20132805 (2020)","journal-title":"IEEE Trans. Med. Imaging"},{"issue":"2","key":"1095_CR7","doi-asserted-by":"publisher","first-page":"1081","DOI":"10.1007\/s40123-023-00651-x","volume":"12","author":"T Wang","year":"2023","unstructured":"T. Wang, G. Liao, L. Chen, Y. Zhuang, S. Zhou, Q. Yuan, M. Zhang, Intelligent diagnosis of multiple peripheral retinal lesions in ultra-widefield fundus images based on deep learning. Ophthalmol. Ther. 12(2), 1081\u20131095 (2023)","journal-title":"Ophthalmol. Ther."},{"key":"1095_CR8","doi-asserted-by":"publisher","first-page":"102527","DOI":"10.1016\/j.media.2022.102527","volume":"81","author":"C Jin","year":"2022","unstructured":"C. Jin, J.K. Udupa, L. Zhao, Y. Tong, D. Odhner, G. Pednekar, D.A. Torigian, Object recognition in medical images via anatomy-guided deep learning. Med. Image Anal. 81, 102527 (2022)","journal-title":"Med. Image Anal."},{"key":"1095_CR9","unstructured":"A. Krizhevsky, I. Sutskever, G.E. Hinton, Imagenet classification with deep convolutional neural networks, in Advances in Neural Information Processing Systems, vol. 25 (2012)"},{"key":"1095_CR10","doi-asserted-by":"crossref","unstructured":"K. He, X. Zhang, S. Ren, J. Sun, Deep residual learning for image recognition, in Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition (2016), pp. 770\u2013778","DOI":"10.1109\/CVPR.2016.90"},{"key":"1095_CR11","doi-asserted-by":"crossref","unstructured":"G. Huang, Z. Liu, L Van Der Maaten, K.Q. Weinberger, Densely connected convolutional networks, in Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition (2017), pp. 4700\u20134708","DOI":"10.1109\/CVPR.2017.243"},{"issue":"1","key":"1095_CR12","doi-asserted-by":"publisher","first-page":"176","DOI":"10.1109\/TPAMI.2017.2782687","volume":"41","author":"FC Ghesu","year":"2017","unstructured":"F.C. Ghesu, B. Georgescu, Y. Zheng, S. Grbic, A. Maier, J. Hornegger, D. Comaniciu, Multi-scale deep reinforcement learning for real-time 3D-landmark detection in CT scans. IEEE Trans. Pattern Anal. Mach. Intell. 41(1), 176\u2013189 (2017)","journal-title":"IEEE Trans. Pattern Anal. Mach. Intell."},{"key":"1095_CR13","doi-asserted-by":"publisher","first-page":"113180","DOI":"10.1016\/j.measurement.2023.113180","volume":"218","author":"H Wen","year":"2023","unstructured":"H. Wen, K. Song, L. Huang, H. Wang, J. Wang, Y. Yan, Hierarchical two-stage modal fusion for triple-modality salient object detection. Measurement 218, 113180 (2023)","journal-title":"Measurement"},{"key":"1095_CR14","doi-asserted-by":"publisher","first-page":"15","DOI":"10.1016\/j.egyr.2023.04.057","volume":"9","author":"W Zha","year":"2023","unstructured":"W. Zha, L. Hu, C. Duan, Y. Li, Semi-supervised learning-based satellite remote sensing object detection method for power transmission towers. Energy Rep. 9, 15\u201327 (2023)","journal-title":"Energy Rep."},{"key":"1095_CR15","doi-asserted-by":"publisher","first-page":"2733","DOI":"10.1007\/s10462-021-10061-9","volume":"55","author":"N Le","year":"2022","unstructured":"N. Le, V.S. Rathour, K. Yamazaki, K. Luu, M. Savvides, Deep reinforcement learning in computer vision: a comprehensive survey. Artifi. Intel. Rev. 55, 2733\u20132819 (2022)","journal-title":"Artifi. Intel. Rev."},{"key":"1095_CR16","unstructured":"F. Navarro, A. Sekuboyina, D. Waldmannstetter, J.C. Peeken, S.E. Combs, B.H. Menze, Deep reinforcement learning for organ localization in CT, in Medical Imaging with Deep Learning (2020), pp. 544\u2013554"},{"issue":"5","key":"1095_CR17","doi-asserted-by":"publisher","first-page":"1143","DOI":"10.1007\/s10278-022-00644-5","volume":"35","author":"JN Stember","year":"2022","unstructured":"J.N. Stember, H. Shalu, Deep reinforcement learning with automated label extraction from clinical reports accurately classifies 3D MRI brain volumes. J. Digit. Imaging 35(5), 1143\u20131152 (2022)","journal-title":"J. Digit. Imaging"},{"key":"1095_CR18","doi-asserted-by":"crossref","unstructured":"G. Maicas, G. Carneiro, A.P. Bradley , J.C. Nascimento, I. Reid, Deep reinforcement learning for active breast lesion detection from DCE-MRI, in International Conference on Medical Image Computing and Computer-Assisted Intervention (2017), pp. 665\u2013673","DOI":"10.1007\/978-3-319-66179-7_76"},{"key":"1095_CR19","doi-asserted-by":"crossref","unstructured":"X. Kong, B. Xin, Y. Wang, G. Hua, Collaborative deep reinforcement learning for joint object search, in Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition (2017), pp. 1695\u20131704","DOI":"10.1109\/CVPR.2017.748"},{"issue":"2","key":"1095_CR20","doi-asserted-by":"publisher","first-page":"639","DOI":"10.1007\/s00371-021-02363-4","volume":"39","author":"H Wang","year":"2023","unstructured":"H. Wang, Y. Chen, M. Wu, X. Zhang, Z. Huang, W. Mao, Attentional and adversarial feature mimic for efficient object detection. Vis. Comput. 39(2), 639\u2013650 (2023)","journal-title":"Vis. Comput."},{"key":"1095_CR21","unstructured":"H. Law, J. Deng, Cornernet: detecting objects as paired keypoints, in Proceedings of the European Conference on Computer Vision, pp. 734\u2013750"},{"key":"1095_CR22","unstructured":"X. Zhou, J. Zhuo, P. Krahenbuhl, Bottom-up object detection by grouping extreme and center points, in Proceedings of the IEEE\/CFV Conference on Computer Vision and Pattern Recognition, pp. 850\u2013859"},{"key":"1095_CR23","doi-asserted-by":"crossref","unstructured":"Z. Dong, G. Li, Y. Liao, F. Wang, P. Ren, C. Qian, Centripetalnet: pursuing high-quality keypoint pairs for object detection, in Proceedings of the IEEE\/CFV Conference on Computer Vision and Pattern Recognition (2020), pp. 10519\u201310528","DOI":"10.1109\/CVPR42600.2020.01053"},{"key":"1095_CR24","doi-asserted-by":"publisher","DOI":"10.1016\/j.patcog.2022.109278","volume":"137","author":"Y Song","year":"2023","unstructured":"Y. Song, P. Zhang, W. Huang, Y. Zha, T. You, Y. Zhang, Object detection based on cortex hierarchical activation in border sensitive mechanism and classification-GIou joint representation. Pattern Recognit. 137, 109278 (2023)","journal-title":"Pattern Recognit."},{"key":"1095_CR25","doi-asserted-by":"publisher","first-page":"104471","DOI":"10.1016\/j.imavis.2022.104471","volume":"123","author":"K Tong","year":"2022","unstructured":"K. Tong, Y. Wu, Deep learning-based detection from the perspective of small or tiny objects: a survey. Image Vis. Comput. 123, 104471 (2022)","journal-title":"Image Vis. Comput."},{"issue":"7540","key":"1095_CR26","doi-asserted-by":"publisher","first-page":"529","DOI":"10.1038\/nature14236","volume":"518","author":"V Mnih","year":"2015","unstructured":"V. Mnih, K. Kavukcuoglu, D. Silver, A.A. Rusu, J. Veness, M.G. Bellemare, D. Hassabis, Human-level control through deep reinforcement learning. Nature 518(7540), 529\u2013533 (2015)","journal-title":"Nature"},{"key":"1095_CR27","doi-asserted-by":"publisher","first-page":"279","DOI":"10.1007\/BF00992698","volume":"8","author":"CJ Watkins","year":"1992","unstructured":"C.J. Watkins, P. Dayan, Q-learning. Mach. Learn. 8, 279\u2013292 (1992)","journal-title":"Mach. Learn."},{"key":"1095_CR28","unstructured":"J. Foerster, I.A. Assael, N. De Freitas, S. Whiteson, Learning to communicate with deep multi-agent reinforcement learning, in Advances in Neural Information Processing Systems (2016), p. 29"},{"key":"1095_CR29","doi-asserted-by":"crossref","unstructured":"J.K. Gupta, M. Egorov, M. Kochenderfer, Cooperative multi-agent control using deep reinforcement learning, in International Conference on Autonomous Agents and Multiagent Systems, pp. 66\u201383","DOI":"10.1007\/978-3-319-71682-4_5"},{"key":"1095_CR30","doi-asserted-by":"crossref","unstructured":"A. Vlontzos, A. Alansary, K. Kamnitsas, D. Rueckert, B. Kainz. Multiple landmark detection using multi-agent reinforcement learning, in Medical Image Computing and Computer Assisted Intervention (2019), pp. 262\u2013270","DOI":"10.1007\/978-3-030-32251-9_29"},{"issue":"3","key":"1095_CR31","doi-asserted-by":"publisher","first-page":"791","DOI":"10.1109\/TMI.2015.2496296","volume":"35","author":"Z Tian","year":"2015","unstructured":"Z. Tian, L. Liu, Z. Zhang, B. Fei, Superpixel-based segmentation for 3D prostate MR images. IEEE Trans. Med. Imaging 35(3), 791\u2013801 (2015)","journal-title":"IEEE Trans. Med. Imaging"},{"key":"1095_CR32","doi-asserted-by":"crossref","unstructured":"J.I. Orlando, H. Fu, J.B. Breda, K. Van Keer, D.R. Bathula, A. Diaz-Pinto, Bogunovi? H, Refuge challenge: A unified framework for evaluating automated methods for glaucoma assessment from fundus photographs. Med. Image Anal. 59, 101570 (2020)","DOI":"10.1016\/j.media.2019.101570"},{"issue":"8","key":"1095_CR33","doi-asserted-by":"publisher","first-page":"1885","DOI":"10.1109\/TMI.2019.2894854","volume":"38","author":"X Xu","year":"2019","unstructured":"X. Xu, F. Zhou, B. Liu, D. Fu, X. Bai, Efficient multiple organ localization in CT image using 3D region proposal network. IEEE Trans. Med. Imaging 38(8), 1885\u20131898 (2019)","journal-title":"IEEE Trans. Med. Imaging"},{"key":"1095_CR34","doi-asserted-by":"crossref","unstructured":"G.E. Humpire-Mamani, A.A.A., Setio B. Van Ginneken, C. Jacobs, Efficient organ localization using multi-label convolutional neural networks in thorax-abdomen CT scans. Phys. Med. Biol. 63(8), 085003 (2018)","DOI":"10.1088\/1361-6560\/aab4b3"},{"key":"1095_CR35","unstructured":"S. Ren, K. He, R. Girshick, J. Sun, Faster R-CNN: towards real-time object detection with region proposal networks, in Advances in Neural Information Processing Systems, vol. 28 (2015)"},{"key":"1095_CR36","doi-asserted-by":"crossref","unstructured":"W. Liu, D. Anguelov, D. Erhan, C. Szegedy, S. Reed, C.Y. Fu, A.C. Berg, Ssd: single shot multibox detector, in European Conference on Computer Vision (2016), pp. 21\u201337","DOI":"10.1007\/978-3-319-46448-0_2"},{"key":"1095_CR37","unstructured":"J. Redmon, A. Farhadi, Yolov3: An incremental improvement (2018). arXiv preprint arXiv:1804.02767"},{"key":"1095_CR38","unstructured":"X. Zhu, W. Su, L. Lu, B. Li, X. Wang, J. Dai, Deformable detr: deformable transformers for end-to-end object detection (2020). arXiv preprint arXiv:2010.04159"},{"key":"1095_CR39","doi-asserted-by":"publisher","first-page":"5011","DOI":"10.1007\/s12652-020-01905-3","volume":"13","author":"Z Tian","year":"2022","unstructured":"Z. Tian, X. Si, Y. Zheng, Z. Chen, X. Li, Multi-step medical image segmentation based on reinforcement learning. J. Ambient Intell. Hum. Comput. 13, 5011\u20135022 (2022)","journal-title":"J. Ambient Intell. Hum. Comput."},{"key":"1095_CR40","unstructured":"A. Bochkovskiy, C.Y. Wang, H.Y.M. Liao, Yolov4: Optimal speed and accuracy of object detection (2020). arXiv preprint arXiv:2004.10934"}],"container-title":["EURASIP Journal on Advances in Signal Processing"],"original-title":[],"language":"en","link":[{"URL":"https:\/\/link.springer.com\/content\/pdf\/10.1186\/s13634-023-01095-y.pdf","content-type":"application\/pdf","content-version":"vor","intended-application":"text-mining"},{"URL":"https:\/\/link.springer.com\/article\/10.1186\/s13634-023-01095-y\/fulltext.html","content-type":"text\/html","content-version":"vor","intended-application":"text-mining"},{"URL":"https:\/\/link.springer.com\/content\/pdf\/10.1186\/s13634-023-01095-y.pdf","content-type":"application\/pdf","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2023,12,21]],"date-time":"2023-12-21T15:03:46Z","timestamp":1703171026000},"score":1,"resource":{"primary":{"URL":"https:\/\/asp-eurasipjournals.springeropen.com\/articles\/10.1186\/s13634-023-01095-y"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2023,12,21]]},"references-count":40,"journal-issue":{"issue":"1","published-online":{"date-parts":[[2023,12]]}},"alternative-id":["1095"],"URL":"https:\/\/doi.org\/10.1186\/s13634-023-01095-y","relation":{},"ISSN":["1687-6180"],"issn-type":[{"value":"1687-6180","type":"electronic"}],"subject":[],"published":{"date-parts":[[2023,12,21]]},"assertion":[{"value":"14 September 2023","order":1,"name":"received","label":"Received","group":{"name":"ArticleHistory","label":"Article History"}},{"value":"10 December 2023","order":2,"name":"accepted","label":"Accepted","group":{"name":"ArticleHistory","label":"Article History"}},{"value":"21 December 2023","order":3,"name":"first_online","label":"First Online","group":{"name":"ArticleHistory","label":"Article History"}},{"order":1,"name":"Ethics","group":{"name":"EthicsHeading","label":"Declarations"}},{"value":"Not applicable.","order":2,"name":"Ethics","group":{"name":"EthicsHeading","label":"Ethics approval and consent to participate"}},{"value":"Not applicable.","order":3,"name":"Ethics","group":{"name":"EthicsHeading","label":"Consent for publication"}},{"value":"The authors declare that they have no competing interests.","order":4,"name":"Ethics","group":{"name":"EthicsHeading","label":"Competing interests"}}],"article-number":"132"}}