{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2026,7,7]],"date-time":"2026-07-07T15:45:15Z","timestamp":1783439115360,"version":"3.54.6"},"reference-count":33,"publisher":"MDPI AG","issue":"7","license":[{"start":{"date-parts":[[2023,3,23]],"date-time":"2023-03-23T00:00:00Z","timestamp":1679529600000},"content-version":"vor","delay-in-days":0,"URL":"https:\/\/creativecommons.org\/licenses\/by\/4.0\/"}],"funder":[{"DOI":"10.13039\/501100001809","name":"National Natural Science Foundation of China","doi-asserted-by":"publisher","award":["62273294"],"award-info":[{"award-number":["62273294"]}],"id":[{"id":"10.13039\/501100001809","id-type":"DOI","asserted-by":"publisher"}]},{"DOI":"10.13039\/501100001809","name":"National Natural Science Foundation of China","doi-asserted-by":"publisher","award":["62033011"],"award-info":[{"award-number":["62033011"]}],"id":[{"id":"10.13039\/501100001809","id-type":"DOI","asserted-by":"publisher"}]},{"DOI":"10.13039\/501100001809","name":"National Natural Science Foundation of China","doi-asserted-by":"publisher","award":["ZD2022104"],"award-info":[{"award-number":["ZD2022104"]}],"id":[{"id":"10.13039\/501100001809","id-type":"DOI","asserted-by":"publisher"}]},{"name":"Science and Technology Project of Hebei Education Department","award":["62273294"],"award-info":[{"award-number":["62273294"]}]},{"name":"Science and Technology Project of Hebei Education Department","award":["62033011"],"award-info":[{"award-number":["62033011"]}]},{"name":"Science and Technology Project of Hebei Education Department","award":["ZD2022104"],"award-info":[{"award-number":["ZD2022104"]}]},{"name":"Introduced Overseas Students of Hebei Province","award":["62273294"],"award-info":[{"award-number":["62273294"]}]},{"name":"Introduced Overseas Students of Hebei Province","award":["62033011"],"award-info":[{"award-number":["62033011"]}]},{"name":"Introduced Overseas Students of Hebei Province","award":["ZD2022104"],"award-info":[{"award-number":["ZD2022104"]}]}],"content-domain":{"domain":[],"crossmark-restriction":false},"short-container-title":["Sensors"],"abstract":"<jats:p>Underwater target detection techniques have been extensively applied to underwater vehicles for marine surveillance, aquaculture, and rescue applications. However, due to complex underwater environments and insufficient training samples, the existing underwater target recognition algorithm accuracy is still unsatisfactory. A long-term effort is essential to improving underwater target detection accuracy. To achieve this goal, in this work, we propose a modified YOLOv5s network, called YOLOv5s-CA network, by embedding a Coordinate Attention (CA) module and a Squeeze-and-Excitation (SE) module, aiming to concentrate more computing power on the target to improve detection accuracy. Based on the existing YOLOv5s network, the number of bottlenecks in the first C3 module was increased from one to three to improve the performance of shallow feature extraction. The CA module was embedded into the C3 modules to improve the attention power focused on the target. The SE layer was added to the output of the C3 modules to strengthen model attention. Experiments on the data of the 2019 China Underwater Robot Competition were conducted, and the results demonstrate that the mean Average Precision (mAP) of the modified YOLOv5s network was increased by 2.4%.<\/jats:p>","DOI":"10.3390\/s23073367","type":"journal-article","created":{"date-parts":[[2023,3,23]],"date-time":"2023-03-23T02:35:26Z","timestamp":1679538926000},"page":"3367","update-policy":"https:\/\/doi.org\/10.3390\/mdpi_crossmark_policy","source":"Crossref","is-referenced-by-count":75,"title":["YOLOv5s-CA: A Modified YOLOv5s Network with Coordinate Attention for Underwater Target Detection"],"prefix":"10.3390","volume":"23","author":[{"given":"Ge","family":"Wen","sequence":"first","affiliation":[{"name":"School of Electrical Engineering, Yanshan University, Qinhuangdao 066004, China"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"ORCID":"https:\/\/orcid.org\/0000-0002-2934-3803","authenticated-orcid":false,"given":"Shaobao","family":"Li","sequence":"additional","affiliation":[{"name":"School of Electrical Engineering, Yanshan University, Qinhuangdao 066004, China"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Fucai","family":"Liu","sequence":"additional","affiliation":[{"name":"School of Electrical Engineering, Yanshan University, Qinhuangdao 066004, China"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"ORCID":"https:\/\/orcid.org\/0000-0003-0404-9533","authenticated-orcid":false,"given":"Xiaoyuan","family":"Luo","sequence":"additional","affiliation":[{"name":"School of Electrical Engineering, Yanshan University, Qinhuangdao 066004, China"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Meng-Joo","family":"Er","sequence":"additional","affiliation":[{"name":"Institute of Artificial Intelligence and Marine Robotics, College of Marine Electrical Engineering, Dalian Maritime University, Dalian 116026, China"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"ORCID":"https:\/\/orcid.org\/0000-0002-2037-8348","authenticated-orcid":false,"given":"Mufti","family":"Mahmud","sequence":"additional","affiliation":[{"name":"Department of Computer Science, Computing and Informatics Research Centre, Medical Technologies Innovation Facility of Nottingham Trent University, Nottingham NG11 8NS, UK"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Tao","family":"Wu","sequence":"additional","affiliation":[{"name":"Department of Frontier & Innovation Research, Wuhan Second Ship Design & Research Institute, Wuhan 430205, China"}],"role":[{"vocabulary":"crossref","role":"author"}]}],"member":"1968","published-online":{"date-parts":[[2023,3,23]]},"reference":[{"key":"ref_1","doi-asserted-by":"crossref","first-page":"108159","DOI":"10.1016\/j.compeleceng.2022.108159","article-title":"Underwater object detection using collaborative weakly supervision","volume":"102","author":"Cai","year":"2022","journal-title":"Comput. Electr. Eng."},{"key":"ref_2","doi-asserted-by":"crossref","first-page":"49","DOI":"10.1016\/j.neucom.2015.10.122","article-title":"DeepFish: Accurate underwater live fish recognition with a deep architecture","volume":"187","author":"Qin","year":"2016","journal-title":"Neurocomputing"},{"key":"ref_3","doi-asserted-by":"crossref","first-page":"39273","DOI":"10.1109\/ACCESS.2020.2976121","article-title":"Multi-AUV collaborative target recognition based on transfer-reinforcement learning","volume":"8","author":"Cai","year":"2020","journal-title":"IEEE Access"},{"key":"ref_4","doi-asserted-by":"crossref","first-page":"436","DOI":"10.1038\/nature14539","article-title":"Deep learning","volume":"521","author":"LeCun","year":"2015","journal-title":"Nature"},{"key":"ref_5","doi-asserted-by":"crossref","first-page":"971","DOI":"10.1109\/TPAMI.2002.1017623","article-title":"Multiresolution gray-scale and rotation invariant texture classification with local binary patterns","volume":"24","author":"Ojala","year":"2002","journal-title":"IEEE Trans. Pattern Anal. Mach. Intell."},{"key":"ref_6","doi-asserted-by":"crossref","first-page":"91","DOI":"10.1023\/B:VISI.0000029664.99615.94","article-title":"Distinctive image features from scale-invariant keypoints","volume":"60","author":"Lowe","year":"2004","journal-title":"Int. J. Comput. Vis."},{"key":"ref_7","unstructured":"Dalal, N., and Triggs, B. (2005, January 20\u201326). Histograms of oriented gradients for human detection. Proceedings of the 2005 IEEE Computer Society Conference on Computer Vision and Pattern Recognition (CVPR\u201905), San Diego, CA, USA."},{"key":"ref_8","doi-asserted-by":"crossref","unstructured":"Villon, S., Chaumont, M., Subsol, G., Vill\u00e9ger, S., Claverie, T., and Mouillot, D. (2016, January 24\u201327). Coral reef fish detection and recognition in underwater videos by supervised machine learning: Comparison between Deep Learning and HOG+ SVM methods. Proceedings of the International Conference on Advanced Concepts for Intelligent Vision Systems, Lecce, Italy.","DOI":"10.1007\/978-3-319-48680-2_15"},{"key":"ref_9","doi-asserted-by":"crossref","unstructured":"Girshick, R. (2015, January 7\u201313). Fast r-cnn. Proceedings of the IEEE International Conference on Computer Vision, Santiago, Chile.","DOI":"10.1109\/ICCV.2015.169"},{"key":"ref_10","doi-asserted-by":"crossref","unstructured":"He, K., Gkioxari, G., Doll\u00e1r, P., and Girshick, R. (2017, January 22\u201329). Mask r-cnn. Proceedings of the IEEE International Conference on Computer Vision, Venice, Italy.","DOI":"10.1109\/ICCV.2017.322"},{"key":"ref_11","first-page":"379","article-title":"R-fcn: Object detection via region-based fully convolutional networks","volume":"29","author":"Dai","year":"2016","journal-title":"Adv. Neural Inf. Process. Syst."},{"key":"ref_12","doi-asserted-by":"crossref","first-page":"104190","DOI":"10.1016\/j.engappai.2021.104190","article-title":"Underwater target detection based on faster r-cnn and adversarial occlusion network","volume":"100","author":"Zeng","year":"2021","journal-title":"Eng. Appl. Artif. Intell."},{"key":"ref_13","doi-asserted-by":"crossref","first-page":"172848","DOI":"10.1109\/ACCESS.2020.3025617","article-title":"Integrate MSRCR and mask R-CNN to recognize underwater creatures on small sample datasets","volume":"8","author":"Song","year":"2020","journal-title":"IEEE Access"},{"key":"ref_14","doi-asserted-by":"crossref","unstructured":"Redmon, J., Divvala, S., Girshick, R., and Farhadi, A. (2016, January 27\u201330). You only look once: Unified, real-time object detection. Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition, Las Vegas, NV, USA.","DOI":"10.1109\/CVPR.2016.91"},{"key":"ref_15","doi-asserted-by":"crossref","unstructured":"Redmon, J., and Farhadi, A. (2017, January 21\u201326). YOLO9000: Better, faster, stronger. Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition, Honolulu, HI, USA.","DOI":"10.1109\/CVPR.2017.690"},{"key":"ref_16","unstructured":"Redmon, J., and Farhadi, A. (2018). Yolov3: An incremental improvement. arXiv."},{"key":"ref_17","unstructured":"Bochkovskiy, A., Wang, C.Y., and Liao, H.Y.M. (2020). Yolov4: Optimal speed and accuracy of object detection. arXiv."},{"key":"ref_18","doi-asserted-by":"crossref","unstructured":"Liu, W., Anguelov, D., Erhan, D., Szegedy, C., Reed, S., Fu, C.Y., and Berg, A.C. (2016, January 11\u201314). Ssd: Single shot multibox detector. Proceedings of the European Conference on Computer Vision, Amsterdam, The Netherlands.","DOI":"10.1007\/978-3-319-46448-0_2"},{"key":"ref_19","doi-asserted-by":"crossref","unstructured":"Chen, L., Zheng, M., Duan, S., Luo, W., and Yao, L. (2021). Underwater Target Recognition Based on Improved YOLOv4 Neural Network. Electronics, 10.","DOI":"10.3390\/electronics10141634"},{"key":"ref_20","doi-asserted-by":"crossref","unstructured":"Yao, Y., Qiu, Z., and Zhong, M. (2019, January 20\u201322). Application of improved MobileNet-SSD on underwater sea cucumber detection robot. Proceedings of the 2019 IEEE 4th Advanced Information Technology, Electronic and Automation Control Conference (IAEAC), Chengdu, China.","DOI":"10.1109\/IAEAC47372.2019.8997970"},{"key":"ref_21","doi-asserted-by":"crossref","first-page":"141861","DOI":"10.1109\/ACCESS.2021.3120870","article-title":"YOLO-FIRI: Improved YOLOv5 for Infrared Image Object Detection","volume":"9","author":"Li","year":"2021","journal-title":"IEEE Access"},{"key":"ref_22","unstructured":"Xu, K., Ba, J., Kiros, R., Cho, K., Courville, A., Salakhudinov, R., Zemel, R., and Bengio, Y. (2015, January 7\u20139). Show, attend and tell: Neural image caption generation with visual attention. Proceedings of the International Conference on Machine Learning, PMLR, Lille, France."},{"key":"ref_23","unstructured":"Tsotsos, J.K. (2021). A Computational Perspective on Visual Attention, MIT Press."},{"key":"ref_24","unstructured":"Mnih, V., Heess, N., and Graves, A. (2014). Recurrent models of visual attention. Adv. Neural Inf. Process. Syst., 27."},{"key":"ref_25","doi-asserted-by":"crossref","unstructured":"Liu, J.J., Hou, Q., Cheng, M.M., Wang, C., and Feng, J. (2020, January 13\u201319). Improving convolutional networks with self-calibrated convolutions. Proceedings of the IEEE\/CVF Conference on Computer Vision and Pattern Recognition, Seattle, WA, USA.","DOI":"10.1109\/CVPR42600.2020.01011"},{"key":"ref_26","doi-asserted-by":"crossref","unstructured":"Bello, I., Zoph, B., Vaswani, A., Shlens, J., and Le, Q.V. (2019, January 27\u201328). Attention augmented convolutional networks. Proceedings of the IEEE\/CVF International Conference on Computer Vision, Seoul, Republic of Korea.","DOI":"10.1109\/ICCV.2019.00338"},{"key":"ref_27","doi-asserted-by":"crossref","unstructured":"Fu, J., Liu, J., Tian, H., Li, Y., Bao, Y., Fang, Z., and Lu, H. (2019, January 15\u201320). Dual attention network for scene segmentation. Proceedings of the IEEE\/CVF Conference on Computer Vision and Pattern Recognition, Long Beach, CA, USA.","DOI":"10.1109\/CVPR.2019.00326"},{"key":"ref_28","doi-asserted-by":"crossref","unstructured":"Shen, Z., and Nguyen, C. (December, January 29). Temporal 3D RetinaNet for fish detection. Proceedings of the 2020 Digital Image Computing: Techniques and Applications (DICTA), Melbourne, Australia.","DOI":"10.1109\/DICTA51227.2020.9363372"},{"key":"ref_29","doi-asserted-by":"crossref","unstructured":"Li, X., Wang, W., Hu, X., and Yang, J. (2019, January 15\u201320). Selective kernel networks. Proceedings of the IEEE\/CVF Conference on Computer Vision and Pattern Recognition, Long Beach, CA, USA.","DOI":"10.1109\/CVPR.2019.00060"},{"key":"ref_30","doi-asserted-by":"crossref","unstructured":"Woo, S., Park, J., Lee, J.Y., and Kweon, I.S. (2018, January 8\u201314). Cbam: Convolutional block attention module. Proceedings of the European Conference on Computer Vision (ECCV), Munich, Germany.","DOI":"10.1007\/978-3-030-01234-2_1"},{"key":"ref_31","doi-asserted-by":"crossref","unstructured":"Misra, D., Nalamada, T., Arasanipalai, A.U., and Hou, Q. (2021, January 3\u20138). Rotate to attend: Convolutional triplet attention module. Proceedings of the IEEE\/CVF Winter Conference on Applications of Computer Vision, Waikoloa, HI, USA.","DOI":"10.1109\/WACV48630.2021.00318"},{"key":"ref_32","doi-asserted-by":"crossref","unstructured":"Hu, J., Shen, L., and Sun, G. (2018, January 18\u201323). Squeeze-and-excitation networks. Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition, Salt Lake City, UT, USA.","DOI":"10.1109\/CVPR.2018.00745"},{"key":"ref_33","doi-asserted-by":"crossref","unstructured":"Hou, Q., Zhou, D., and Feng, J. (2021, January 20\u201325). Coordinate attention for efficient mobile network design. Proceedings of the IEEE\/CVF Conference on Computer Vision and Pattern Recognition, Nashville, TN, USA.","DOI":"10.1109\/CVPR46437.2021.01350"}],"container-title":["Sensors"],"original-title":[],"language":"en","link":[{"URL":"https:\/\/www.mdpi.com\/1424-8220\/23\/7\/3367\/pdf","content-type":"unspecified","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2025,10,10]],"date-time":"2025-10-10T19:01:01Z","timestamp":1760122861000},"score":1,"resource":{"primary":{"URL":"https:\/\/www.mdpi.com\/1424-8220\/23\/7\/3367"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2023,3,23]]},"references-count":33,"journal-issue":{"issue":"7","published-online":{"date-parts":[[2023,4]]}},"alternative-id":["s23073367"],"URL":"https:\/\/doi.org\/10.3390\/s23073367","relation":{},"ISSN":["1424-8220"],"issn-type":[{"value":"1424-8220","type":"electronic"}],"subject":[],"published":{"date-parts":[[2023,3,23]]}}}