{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2026,7,16]],"date-time":"2026-07-16T11:13:15Z","timestamp":1784200395951,"version":"3.55.0"},"reference-count":58,"publisher":"Association for Computing Machinery (ACM)","issue":"8","content-domain":{"domain":["dl.acm.org"],"crossmark-restriction":true},"short-container-title":["ACM Trans. Multimedia Comput. Commun. Appl."],"published-print":{"date-parts":[[2025,8,31]]},"abstract":"<jats:p>Underwater visual object tracking is crucial for marine resource exploration and military security. However, due to the effect of insufficient light and turbid background in underwater scenes, efficient and accurate target tracking cannot be realized on underwater edge devices with limited computing resources. To address this problem, we design an underwater object tracking network, namely DBSF, for edge computing devices based on sparse confidence feature learning guided by differential boundary attention. Specifically, we propose a differential boundary attention distribution model to compute the object edge distribution state to enhance the accurate perception of the underwater object edge structure. Then, the differential boundary attention-guided object tracking network learns to perceive the highly discriminative sparse features on the object structure, and computes the object sparse confidence matrix, which reduces the constraints of the edge devices with limited computational resources and ensures the tracking performance. Extensive experiments demonstrate that the DBSF network achieves accurate underwater target recognition and outperforms related advanced methods.<\/jats:p>","DOI":"10.1145\/3689824","type":"journal-article","created":{"date-parts":[[2024,8,24]],"date-time":"2024-08-24T12:13:18Z","timestamp":1724501598000},"page":"1-17","update-policy":"https:\/\/doi.org\/10.1145\/crossmark-policy","source":"Crossref","is-referenced-by-count":5,"title":["Boundary Attention-Guided Sparse Feature Learning for Underwater Object Tracking in Edge Computing"],"prefix":"10.1145","volume":"21","author":[{"ORCID":"https:\/\/orcid.org\/0009-0009-9400-6107","authenticated-orcid":false,"given":"Hongyi","family":"Qiu","sequence":"first","affiliation":[{"name":"School of Mathematics and Statistics, Changchun University of Science and Technology, Changchun, China"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"ORCID":"https:\/\/orcid.org\/0000-0002-7548-9379","authenticated-orcid":false,"given":"Ning","family":"Li","sequence":"additional","affiliation":[{"name":"Institute of Artificial Intelligence, Guangzhou University, Guangzhou, China"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"ORCID":"https:\/\/orcid.org\/0009-0001-8190-6239","authenticated-orcid":false,"given":"Pengfei","family":"Li","sequence":"additional","affiliation":[{"name":"School of Civil Engineering and Architecture, Zhengzhou University of Aeronautics, Zhengzhou, China"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"ORCID":"https:\/\/orcid.org\/0000-0003-3734-7350","authenticated-orcid":false,"given":"Ruitao","family":"Hou","sequence":"additional","affiliation":[{"name":"Institute of Artificial Intelligence, Guangzhou University, Guangzhou, China"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"ORCID":"https:\/\/orcid.org\/0009-0007-8938-7705","authenticated-orcid":false,"given":"Yuting","family":"Zhang","sequence":"additional","affiliation":[{"name":"Institute of Artificial Intelligence, Guangzhou University, Guangzhou, China"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"ORCID":"https:\/\/orcid.org\/0000-0001-6358-2333","authenticated-orcid":false,"given":"Yun","family":"Peng","sequence":"additional","affiliation":[{"name":"Institute of Artificial Intelligence, Guangzhou University, Guangzhou, China"}],"role":[{"vocabulary":"crossref","role":"author"}]}],"member":"320","published-online":{"date-parts":[[2025,8,13]]},"reference":[{"key":"e_1_3_1_2_2","first-page":"3326","volume-title":"Proceedings of the Asian Conference on Computer Vision","author":"Alawode Basit","year":"2022","unstructured":"Basit Alawode, Yuhang Guo, Mehnaz Ummar, Naoufel Werghi, Jorge Dias, Ajmal Mian, and Sajid Javed. 2022. UTB180: A high-quality benchmark for underwater tracking. In Proceedings of the Asian Conference on Computer Vision, 3326\u20133342."},{"key":"e_1_3_1_3_2","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR.2016.156"},{"key":"e_1_3_1_4_2","doi-asserted-by":"crossref","first-page":"850","DOI":"10.1007\/978-3-319-48881-3_56","volume-title":"Proceedings of the Computer Vision\u2013ECCV 2016 Workshops","author":"Bertinetto Luca","year":"2016","unstructured":"Luca Bertinetto, Jack Valmadre, Joao F. Henriques, Andrea Vedaldi, and Philip H. S. Torr. 2016b. Fully-convolutional Siamese networks for object tracking. In Proceedings of the Computer Vision\u2013ECCV 2016 Workshops, Proceedings, Part II 14. Springer, 850\u2013865."},{"key":"e_1_3_1_5_2","doi-asserted-by":"publisher","DOI":"10.1109\/ICCV.2019.00628"},{"key":"e_1_3_1_6_2","first-page":"1571","volume-title":"Proceedings of the IEEE\/CVF Winter Conference on Applications of Computer Vision","author":"Blatter Philippe","year":"2023","unstructured":"Philippe Blatter, Menelaos Kanakis, Martin Danelljan, and Luc Van Gool. 2023. Efficient visual tracking with exemplar transformers. In Proceedings of the IEEE\/CVF Winter Conference on Applications of Computer Vision, 1571\u20131581."},{"key":"e_1_3_1_7_2","doi-asserted-by":"publisher","DOI":"10.1007\/978-3-030-58452-8_13"},{"key":"e_1_3_1_8_2","doi-asserted-by":"publisher","DOI":"10.1109\/TNNLS.2019.2927224"},{"key":"e_1_3_1_9_2","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR52729.2023.01400"},{"key":"e_1_3_1_10_2","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR46437.2021.00803"},{"issue":"4","key":"e_1_3_1_11_2","doi-asserted-by":"crossref","first-page":"3288","DOI":"10.1109\/TIE.2019.2913815","article-title":"Visual tracking via auto-encoder pair correlation filter","volume":"67","author":"Cheng Xu","year":"2019","unstructured":"Xu Cheng, Yifeng Zhang, Lin Zhou, and Yuhui Zheng. 2019. Visual tracking via auto-encoder pair correlation filter. IEEE Transactions on Industrial Electronics 67, 4 (2019), 3288\u20133297.","journal-title":"IEEE Transactions on Industrial Electronics"},{"key":"e_1_3_1_12_2","first-page":"4870","volume-title":"Proceedings of the IEEE\/CVF Winter Conference on Applications of Computer Vision","author":"Chu Peng","year":"2023","unstructured":"Peng Chu, Jiang Wang, Quanzeng You, Haibin Ling, and Zicheng Liu. 2023. TransMOT: Spatial-temporal graph transformer for multiple object tracking. In Proceedings of the IEEE\/CVF Winter Conference on Applications of Computer Vision, 4870\u20134880."},{"key":"e_1_3_1_13_2","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR52688.2022.01324"},{"key":"e_1_3_1_14_2","first-page":"58736","article-title":"MixFormerv2: Efficient fully transformer tracking","volume":"36","author":"Cui Yutao","year":"2023","unstructured":"Yutao Cui, Tianhui Song, Gangshan Wu, and Limin Wang. 2023. MixFormerv2: Efficient fully transformer tracking. Advances in Neural Information Processing Systems 36 (2023), 58736\u201358751.","journal-title":"Advances in Neural Information Processing Systems"},{"key":"e_1_3_1_15_2","doi-asserted-by":"publisher","DOI":"10.1109\/ICCV.2015.490"},{"key":"e_1_3_1_16_2","unstructured":"Alexey Dosovitskiy Lucas Beyer Alexander Kolesnikov Dirk Weissenborn Xiaohua Zhai Thomas Unterthiner Mostafa Dehghani Matthias Minderer Georg Heigold Sylvain Gelly Jakob Uszkoreit and Neil Houlsby. 2020. An image is worth 16x16 words: Transformers for image recognition at scale. arXiv:2010.11929."},{"key":"e_1_3_1_17_2","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR.2019.00552"},{"issue":"1","key":"e_1_3_1_18_2","first-page":"1417","article-title":"Siamese object tracking for unmanned aerial vehicle: A review and comprehensive analysis","volume":"56","author":"Fu Changhong","year":"2023","unstructured":"Changhong Fu, Kunhan Lu, Guangze Zheng, Junjie Ye, Ziang Cao, Bowen Li, and Geng Lu. 2023. Siamese object tracking for unmanned aerial vehicle: A review and comprehensive analysis. Artificial Intelligence Review 56, Suppl 1 (2023), 1417\u20131477.","journal-title":"Artificial Intelligence Review"},{"key":"e_1_3_1_19_2","doi-asserted-by":"publisher","DOI":"10.1007\/978-3-031-20047-2_9"},{"key":"e_1_3_1_20_2","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR42600.2020.00630"},{"issue":"1","key":"e_1_3_1_21_2","first-page":"155","article-title":"Adaptive discriminative deep correlation filter for visual object tracking","volume":"30","author":"Han Zhenjun","year":"2018","unstructured":"Zhenjun Han, Pan Wang, and Qixiang Ye. 2018. Adaptive discriminative deep correlation filter for visual object tracking. IEEE Transactions on Circuits and Systems for Video Technology 30, 1 (2018), 155\u2013166.","journal-title":"IEEE Transactions on Circuits and Systems for Video Technology"},{"key":"e_1_3_1_22_2","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR.2016.90"},{"key":"e_1_3_1_23_2","doi-asserted-by":"publisher","DOI":"10.1109\/TPAMI.2014.2345390"},{"key":"e_1_3_1_24_2","doi-asserted-by":"publisher","DOI":"10.1109\/MNET.011.2000684"},{"key":"e_1_3_1_25_2","doi-asserted-by":"publisher","DOI":"10.1109\/TPAMI.2019.2957464"},{"key":"e_1_3_1_26_2","doi-asserted-by":"publisher","DOI":"10.5555\/2354409.2354884"},{"key":"e_1_3_1_27_2","first-page":"1","volume-title":"Proceedings of the 2015 4th International Conference on Reliability, Infocom Technologies and Optimization (ICRITO)","author":"Kale Kiran","year":"2015","unstructured":"Kiran Kale, Sushant Pawar, and Pravin Dhulekar. 2015. Moving object tracking using optical flow and motion vector estimation. In Proceedings of the 2015 4th International Conference on Reliability, Infocom Technologies and Optimization (ICRITO),Trends and Future Directions. IEEE, 1\u20136."},{"key":"e_1_3_1_28_2","first-page":"1","volume-title":"Proceedings of the 2019 IEEE International Symposium on Technologies for Homeland Security (HST).","author":"Kezebou Landry","year":"2019","unstructured":"Landry Kezebou, Victor Oludare, Karen Panetta, and Sos S Agaian. 2019. Underwater object tracking benchmark and dataset. In Proceedings of the 2019 IEEE International Symposium on Technologies for Homeland Security (HST). IEEE, 1\u20136."},{"key":"e_1_3_1_29_2","doi-asserted-by":"publisher","DOI":"10.3390\/pr11020312"},{"key":"e_1_3_1_30_2","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR.2019.00441"},{"key":"e_1_3_1_31_2","doi-asserted-by":"publisher","DOI":"10.1007\/s00530-023-01064-3"},{"issue":"5","key":"e_1_3_1_32_2","first-page":"053012","article-title":"UStark: Underwater image domain-adaptive tracker based on Stark","volume":"31","author":"Li Yunfeng","year":"2022","unstructured":"Yunfeng Li, Wei Huo, Zhuoyan Liu, Bo Wang, and Ye Li. 2022. UStark: Underwater image domain-adaptive tracker based on Stark. Journal of Electronic Imaging 31, 5 (2022), 053012\u2013053012.","journal-title":"Journal of Electronic Imaging"},{"key":"e_1_3_1_33_2","doi-asserted-by":"publisher","DOI":"10.1016\/j.oceaneng.2023.115449"},{"key":"e_1_3_1_34_2","doi-asserted-by":"publisher","DOI":"10.1007\/978-3-319-10602-1_48"},{"key":"e_1_3_1_35_2","doi-asserted-by":"publisher","DOI":"10.1007\/s40747-020-00161-4"},{"key":"e_1_3_1_36_2","doi-asserted-by":"publisher","DOI":"10.1109\/ICCV.2015.352"},{"key":"e_1_3_1_37_2","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR52688.2022.00864"},{"issue":"1","key":"e_1_3_1_38_2","doi-asserted-by":"crossref","first-page":"502","DOI":"10.1109\/TNNLS.2021.3097498","article-title":"Robust visual tracking via multitask sparse correlation filters learning","volume":"34","author":"Nai Ke","year":"2023","unstructured":"Ke Nai, Zhiyong Li, Yihui Gan, and Qi Wang. 2023. Robust visual tracking via multitask sparse correlation filters learning. IEEE Transactions on Neural Networks and Learning Systems 34, 1 (2023), 502\u2013515.","journal-title":"IEEE Transactions on Neural Networks and Learning Systems"},{"key":"e_1_3_1_39_2","doi-asserted-by":"publisher","DOI":"10.1109\/JOE.2021.3086907"},{"key":"e_1_3_1_40_2","unstructured":"Mia Gaia Polansky Charles Herrmann Junhwa Hur Deqing Sun Dor Verbin and Todd Zickler. 2024. Boundary attention: Learning to find faint boundaries at any resolution. arXiv:2401.00935."},{"key":"e_1_3_1_41_2","doi-asserted-by":"publisher","DOI":"10.1109\/TIP.2018.2797482"},{"key":"e_1_3_1_42_2","doi-asserted-by":"publisher","DOI":"10.1109\/TII.2019.2946618"},{"key":"e_1_3_1_43_2","doi-asserted-by":"publisher","DOI":"10.1109\/TCSVT.2023.3249468"},{"key":"e_1_3_1_44_2","first-page":"12112","volume-title":"Proceedings of the AAAI Conference on Artificial Intelligence","author":"Vihlman Mikko","year":"2020","unstructured":"Mikko Vihlman and Arto Visala. 2020. Optical flow in deep visual tracking. In Proceedings of the AAAI Conference on Artificial Intelligence, Vol. 34, 12112\u201312119."},{"key":"e_1_3_1_45_2","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR42600.2020.00661"},{"key":"e_1_3_1_46_2","doi-asserted-by":"publisher","DOI":"10.1109\/ICCV.2015.357"},{"key":"e_1_3_1_47_2","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR46437.2021.00162"},{"issue":"10","key":"e_1_3_1_48_2","first-page":"3727","article-title":"Learning low-rank and sparse discriminative correlation filters for coarse-to-fine visual object tracking","volume":"30","author":"Xu Tianyang","year":"2019","unstructured":"Tianyang Xu, Zhen-Hua Feng, Xiao-Jun Wu, and Josef Kittler. 2019. Learning low-rank and sparse discriminative correlation filters for coarse-to-fine visual object tracking. IEEE Transactions on Circuits and Systems for Video Technology 30, 10 (2019), 3727\u20133739.","journal-title":"IEEE Transactions on Circuits and Systems for Video Technology"},{"key":"e_1_3_1_49_2","doi-asserted-by":"publisher","DOI":"10.1109\/ICCV48922.2021.01028"},{"key":"e_1_3_1_50_2","doi-asserted-by":"publisher","DOI":"10.1007\/978-3-031-20047-2_20"},{"key":"e_1_3_1_51_2","doi-asserted-by":"publisher","DOI":"10.1016\/j.ins.2016.05.019"},{"key":"e_1_3_1_52_2","first-page":"1","article-title":"Active learning for deep visual tracking","author":"Yuan Di","year":"2023","unstructured":"Di Yuan, Xiaojun Chang, Qiao Liu, Yi Yang, Dehua Wang, Minglei Shu, Zhenyu He, and Guangming Shi. 2023. Active learning for deep visual tracking. IEEE Transactions on Neural Networks and Learning Systems (2023), 1\u201313.","journal-title":"IEEE Transactions on Neural Networks and Learning Systems"},{"issue":"1","key":"e_1_3_1_53_2","first-page":"96","article-title":"Deep position-sensitive tracking","volume":"22","author":"Zha Yufei","year":"2019","unstructured":"Yufei Zha, Tao Ku, Yunqiang Li, and Peng Zhang. 2019. Deep position-sensitive tracking. IEEE Transactions on Multimedia 22, 1 (2019), 96\u2013107.","journal-title":"IEEE Transactions on Multimedia"},{"key":"e_1_3_1_54_2","doi-asserted-by":"publisher","DOI":"10.1016\/j.ins.2023.03.083"},{"key":"e_1_3_1_55_2","doi-asserted-by":"publisher","DOI":"10.1109\/LSP.2023.3238277"},{"key":"e_1_3_1_56_2","doi-asserted-by":"publisher","DOI":"10.1007\/s12652-020-02572-0"},{"key":"e_1_3_1_57_2","doi-asserted-by":"crossref","first-page":"771","DOI":"10.1007\/978-3-030-58589-1_46","volume-title":"Proceedings of the Computer Vision\u2013ECCV 2020: 16th European Conference","author":"Zhang Zhipeng","year":"2020","unstructured":"Zhipeng Zhang, Houwen Peng, Jianlong Fu, Bing Li, and Weiming Hu. 2020. Ocean: Object-aware anchor-free tracking. In Proceedings of the Computer Vision\u2013ECCV 2020: 16th European Conference, Proceedings, Part XXI 16. Springer, 771\u2013787."},{"key":"e_1_3_1_58_2","first-page":"12993","volume-title":"Proceedings of the AAAI Conference on Artificial Intelligence","author":"Zheng Zhaohui","year":"2020","unstructured":"Zhaohui Zheng, Ping Wang, Wei Liu, Jinze Li, Rongguang Ye, and Dongwei Ren. 2020. Distance-IoU loss: Faster and better learning for bounding box regression. In Proceedings of the AAAI Conference on Artificial Intelligence, Vol. 34, 12993\u201313000."},{"key":"e_1_3_1_59_2","doi-asserted-by":"publisher","DOI":"10.5555\/2354409.2354886"}],"container-title":["ACM Transactions on Multimedia Computing, Communications, and Applications"],"original-title":[],"language":"en","link":[{"URL":"https:\/\/dl.acm.org\/doi\/pdf\/10.1145\/3689824","content-type":"unspecified","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2025,8,13]],"date-time":"2025-08-13T12:04:56Z","timestamp":1755086696000},"score":1,"resource":{"primary":{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/3689824"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2025,8,13]]},"references-count":58,"journal-issue":{"issue":"8","published-print":{"date-parts":[[2025,8,31]]}},"alternative-id":["10.1145\/3689824"],"URL":"https:\/\/doi.org\/10.1145\/3689824","relation":{},"ISSN":["1551-6857","1551-6865"],"issn-type":[{"value":"1551-6857","type":"print"},{"value":"1551-6865","type":"electronic"}],"subject":[],"published":{"date-parts":[[2025,8,13]]},"assertion":[{"value":"2024-02-26","order":0,"name":"received","label":"Received","group":{"name":"publication_history","label":"Publication History"}},{"value":"2024-08-16","order":2,"name":"accepted","label":"Accepted","group":{"name":"publication_history","label":"Publication History"}},{"value":"2025-08-13","order":3,"name":"published","label":"Published","group":{"name":"publication_history","label":"Publication History"}}]}}