{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2026,8,15]],"date-time":"2026-08-15T18:13:18Z","timestamp":1786817598681,"version":"3.56.0"},"reference-count":54,"publisher":"MDPI AG","issue":"24","license":[{"start":{"date-parts":[[2023,12,7]],"date-time":"2023-12-07T00:00:00Z","timestamp":1701907200000},"content-version":"vor","delay-in-days":0,"URL":"https:\/\/creativecommons.org\/licenses\/by\/4.0\/"}],"funder":[{"name":"National Natural Science Foundation of China","award":["61873307"],"award-info":[{"award-number":["61873307"]}]},{"name":"National Natural Science Foundation of China","award":["F2020501040"],"award-info":[{"award-number":["F2020501040"]}]},{"name":"National Natural Science Foundation of China","award":["F2021203070"],"award-info":[{"award-number":["F2021203070"]}]},{"name":"National Natural Science Foundation of China","award":["F2022501031"],"award-info":[{"award-number":["F2022501031"]}]},{"name":"National Natural Science Foundation of China","award":["N2123004"],"award-info":[{"award-number":["N2123004"]}]},{"name":"National Natural Science Foundation of China","award":["2022GFZD014"],"award-info":[{"award-number":["2022GFZD014"]}]},{"name":"National Natural Science Foundation of China","award":["206Z1702G"],"award-info":[{"award-number":["206Z1702G"]}]},{"name":"Hebei Natural Science Foundation","award":["61873307"],"award-info":[{"award-number":["61873307"]}]},{"name":"Hebei Natural Science Foundation","award":["F2020501040"],"award-info":[{"award-number":["F2020501040"]}]},{"name":"Hebei Natural Science Foundation","award":["F2021203070"],"award-info":[{"award-number":["F2021203070"]}]},{"name":"Hebei Natural Science Foundation","award":["F2022501031"],"award-info":[{"award-number":["F2022501031"]}]},{"name":"Hebei Natural Science Foundation","award":["N2123004"],"award-info":[{"award-number":["N2123004"]}]},{"name":"Hebei Natural Science Foundation","award":["2022GFZD014"],"award-info":[{"award-number":["2022GFZD014"]}]},{"name":"Hebei Natural Science Foundation","award":["206Z1702G"],"award-info":[{"award-number":["206Z1702G"]}]},{"name":"Fundamental Research Funds for the Central Universities","award":["61873307"],"award-info":[{"award-number":["61873307"]}]},{"name":"Fundamental Research Funds for the Central Universities","award":["F2020501040"],"award-info":[{"award-number":["F2020501040"]}]},{"name":"Fundamental Research Funds for the Central Universities","award":["F2021203070"],"award-info":[{"award-number":["F2021203070"]}]},{"name":"Fundamental Research Funds for the Central Universities","award":["F2022501031"],"award-info":[{"award-number":["F2022501031"]}]},{"name":"Fundamental Research Funds for the Central Universities","award":["N2123004"],"award-info":[{"award-number":["N2123004"]}]},{"name":"Fundamental Research Funds for the Central Universities","award":["2022GFZD014"],"award-info":[{"award-number":["2022GFZD014"]}]},{"name":"Fundamental Research Funds for the Central Universities","award":["206Z1702G"],"award-info":[{"award-number":["206Z1702G"]}]},{"name":"Administration of Central Funds Guiding the Local Science and Technology Development","award":["61873307"],"award-info":[{"award-number":["61873307"]}]},{"name":"Administration of Central Funds Guiding the Local Science and Technology Development","award":["F2020501040"],"award-info":[{"award-number":["F2020501040"]}]},{"name":"Administration of Central Funds Guiding the Local Science and Technology Development","award":["F2021203070"],"award-info":[{"award-number":["F2021203070"]}]},{"name":"Administration of Central Funds Guiding the Local Science and Technology Development","award":["F2022501031"],"award-info":[{"award-number":["F2022501031"]}]},{"name":"Administration of Central Funds Guiding the Local Science and Technology Development","award":["N2123004"],"award-info":[{"award-number":["N2123004"]}]},{"name":"Administration of Central Funds Guiding the Local Science and Technology Development","award":["2022GFZD014"],"award-info":[{"award-number":["2022GFZD014"]}]},{"name":"Administration of Central Funds Guiding the Local Science and Technology Development","award":["206Z1702G"],"award-info":[{"award-number":["206Z1702G"]}]}],"content-domain":{"domain":[],"crossmark-restriction":false},"short-container-title":["Remote Sensing"],"abstract":"<jats:p>Tracking and segmenting small targets in remote sensing videos on edge devices carries significant engineering implications. However, many semi-supervised video object segmentation (S-VOS) methods heavily rely on extensive video random-access memory (VRAM) resources, making deployment on edge devices challenging. Our goal is to develop an edge-deployable S-VOS method that can achieve high-precision tracking and segmentation by selecting a bounding box for the target object. First, a tracker is introduced to pinpoint the position of the tracked object in different frames, thereby eliminating the need to save the results of the split as other S-VOS methods do, thus avoiding an increase in VRAM usage. Second, we use two key lightweight components, correlation filters (CFs) and the Mobile Segment Anything Model (MobileSAM), to ensure the inference speed of our model. Third, a mask diffusion module is proposed that improves the accuracy and robustness of segmentation without increasing VRAM usage. We use our self-built dataset containing airplanes and vehicles to evaluate our method. The results show that on the GTX 1080 Ti, our model achieves a J&amp;F score of 66.4% under the condition that the VRAM usage is less than 500 MB, while maintaining a processing speed of 12 frames per second (FPS). The model we propose exhibits good performance in tracking and segmenting small targets on edge devices, providing a solution for fields such as aircraft monitoring and vehicle tracking that require executing S-VOS tasks on edge devices.<\/jats:p>","DOI":"10.3390\/rs15245665","type":"journal-article","created":{"date-parts":[[2023,12,8]],"date-time":"2023-12-08T03:03:33Z","timestamp":1702004613000},"page":"5665","update-policy":"https:\/\/doi.org\/10.3390\/mdpi_crossmark_policy","source":"Crossref","is-referenced-by-count":18,"title":["MobileSAM-Track: Lightweight One-Shot Tracking and Segmentation of Small Objects on Edge Devices"],"prefix":"10.3390","volume":"15","author":[{"given":"Yehui","family":"Liu","sequence":"first","affiliation":[{"name":"School of Control Engineering, Northeastern University at Qinhuangdao, Qinhuangdao 066000, China"},{"name":"Hebei Key Laboratory of Micro-Nano Precision Optical Sensing and Measurement Technology, Northeastern University at Qinhuangdao, Qinhuangdao 066000, China"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"ORCID":"https:\/\/orcid.org\/0000-0002-4519-7404","authenticated-orcid":false,"given":"Yuliang","family":"Zhao","sequence":"additional","affiliation":[{"name":"School of Control Engineering, Northeastern University at Qinhuangdao, Qinhuangdao 066000, China"},{"name":"Hebei Key Laboratory of Micro-Nano Precision Optical Sensing and Measurement Technology, Northeastern University at Qinhuangdao, Qinhuangdao 066000, China"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Xinyue","family":"Zhang","sequence":"additional","affiliation":[{"name":"School of Control Engineering, Northeastern University at Qinhuangdao, Qinhuangdao 066000, China"},{"name":"Hebei Key Laboratory of Micro-Nano Precision Optical Sensing and Measurement Technology, Northeastern University at Qinhuangdao, Qinhuangdao 066000, China"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Xiaoai","family":"Wang","sequence":"additional","affiliation":[{"name":"School of Control Engineering, Northeastern University at Qinhuangdao, Qinhuangdao 066000, China"},{"name":"Hebei Key Laboratory of Micro-Nano Precision Optical Sensing and Measurement Technology, Northeastern University at Qinhuangdao, Qinhuangdao 066000, China"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"ORCID":"https:\/\/orcid.org\/0000-0002-1919-3063","authenticated-orcid":false,"given":"Chao","family":"Lian","sequence":"additional","affiliation":[{"name":"School of Control Engineering, Northeastern University at Qinhuangdao, Qinhuangdao 066000, China"},{"name":"Hebei Key Laboratory of Micro-Nano Precision Optical Sensing and Measurement Technology, Northeastern University at Qinhuangdao, Qinhuangdao 066000, China"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Jian","family":"Li","sequence":"additional","affiliation":[{"name":"School of Control Engineering, Northeastern University at Qinhuangdao, Qinhuangdao 066000, China"},{"name":"Hebei Key Laboratory of Micro-Nano Precision Optical Sensing and Measurement Technology, Northeastern University at Qinhuangdao, Qinhuangdao 066000, China"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Peng","family":"Shan","sequence":"additional","affiliation":[{"name":"School of Control Engineering, Northeastern University at Qinhuangdao, Qinhuangdao 066000, China"},{"name":"Hebei Key Laboratory of Micro-Nano Precision Optical Sensing and Measurement Technology, Northeastern University at Qinhuangdao, Qinhuangdao 066000, China"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Changzeng","family":"Fu","sequence":"additional","affiliation":[{"name":"School of Control Engineering, Northeastern University at Qinhuangdao, Qinhuangdao 066000, China"},{"name":"Hebei Key Laboratory of Micro-Nano Precision Optical Sensing and Measurement Technology, Northeastern University at Qinhuangdao, Qinhuangdao 066000, China"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Xiaoyong","family":"Lyu","sequence":"additional","affiliation":[{"name":"School of Control Engineering, Northeastern University at Qinhuangdao, Qinhuangdao 066000, China"},{"name":"Hebei Key Laboratory of Micro-Nano Precision Optical Sensing and Measurement Technology, Northeastern University at Qinhuangdao, Qinhuangdao 066000, China"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Lianjiang","family":"Li","sequence":"additional","affiliation":[{"name":"School of Control Engineering, Northeastern University at Qinhuangdao, Qinhuangdao 066000, China"},{"name":"Hebei Key Laboratory of Micro-Nano Precision Optical Sensing and Measurement Technology, Northeastern University at Qinhuangdao, Qinhuangdao 066000, China"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Qiang","family":"Fu","sequence":"additional","affiliation":[{"name":"Shijiazhuang School, Army Engineering University of PLA, Shijiazhuang 050003, China"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Wen Jung","family":"Li","sequence":"additional","affiliation":[{"name":"Department of Mechanical Engineering, City University of Hong Kong, Hong Kong 999077, China"}],"role":[{"vocabulary":"crossref","role":"author"}]}],"member":"1968","published-online":{"date-parts":[[2023,12,7]]},"reference":[{"key":"ref_1","doi-asserted-by":"crossref","first-page":"5184","DOI":"10.1109\/ACCESS.2022.3140876","article-title":"Aircraft Target Detection in Remote Sensing Images Based on Improved YOLOv5","volume":"10","author":"Luo","year":"2022","journal-title":"IEEE Access"},{"key":"ref_2","first-page":"4685644","article-title":"Aircraft Detection for Remote Sensing Images Based on Deep Convolutional Neural Networks","volume":"2021","author":"Zhou","year":"2021","journal-title":"J. Electr. Comput. Eng."},{"key":"ref_3","doi-asserted-by":"crossref","unstructured":"Li, Y., Zhao, J., Zhang, S., and Tan, W. (2018, January 20\u201321). Aircraft Detection in Remote Sensing Images Based on Deep Convolutional Neural Network. Proceedings of the 2018 IEEE 3rd International Conference on Cloud Computing and Internet of Things (CCIOT) Aircraft, Dalian, China.","DOI":"10.1109\/CCIOT45285.2018.9032512"},{"key":"ref_4","doi-asserted-by":"crossref","unstructured":"Wu, S., Zhang, K., Li, S., and Yan, J. (2020). Learning to Track Aircraft in Infrared Imagery. Remote Sens., 12.","DOI":"10.3390\/rs12233995"},{"key":"ref_5","doi-asserted-by":"crossref","unstructured":"Oh, S.W., Lee, J.-Y., Xu, N., and Kim, S.J. (November, January 27). Video Object Segmentation Using Space-Time Memory Networks. Proceedings of the 2019 IEEE\/CVF International Conference on Computer Vision (ICCV), Seoul, Republic of Korea.","DOI":"10.1109\/ICCV.2019.00932"},{"key":"ref_6","unstructured":"Cheng, H.K., Tai, Y.-W., and Tang, C.-K. (2021, January 9). Rethinking Space-Time Networks with Improved Memory Coverage for Efficient Video Object Segmentation. Proceedings of the Advances in Neural Information Processing Systems, New Orleans, LA, USA."},{"key":"ref_7","doi-asserted-by":"crossref","unstructured":"Wang, H., Jiang, X., Ren, H., Hu, Y., and Bai, S. (2021, January 20\u201325). SwiftNet: Real-Time Video Object Segmentation. Proceedings of the IEEE Computer Society Conference on Computer Vision and Pattern Recognition, Nashville, TN, USA.","DOI":"10.1109\/CVPR46437.2021.00135"},{"key":"ref_8","doi-asserted-by":"crossref","unstructured":"Kirillov, A., Mintun, E., Ravi, N., Mao, H., Rolland, C., Gustafson, L., Xiao, T., Whitehead, S., Berg, A.C., and Lo, W.-Y. (2023, January 2\u20136). Segment Anything. Proceedings of the IEEE\/CVF International Conference on Computer Vision (ICCV), Paris, France.","DOI":"10.1109\/ICCV51070.2023.00371"},{"key":"ref_9","unstructured":"Chen, K., Liu, C., Chen, H., Zhang, H., Li, W., Zou, Z., and Shi, Z. (2023). RSPrompter: Learning to Prompt for Remote Sensing Instance Segmentation Based on Visual Foundation Model. arXiv."},{"key":"ref_10","doi-asserted-by":"crossref","unstructured":"Wang, Y., Zhao, Y., and Petzold, L. (2023). An Empirical Study on the Robustness of the Segment Anything Model (SAM). arXiv.","DOI":"10.2139\/ssrn.4476683"},{"key":"ref_11","doi-asserted-by":"crossref","unstructured":"Huang, Y., Yang, X., Liu, L., Zhou, H., Chang, A., Zhou, X., Chen, R., Yu, J., Chen, J., and Chen, C. (2023). Segment Anything Model for Medical Images?. arXiv.","DOI":"10.1016\/j.media.2023.103061"},{"key":"ref_12","doi-asserted-by":"crossref","unstructured":"Caelles, S., Maninis, K.K., Pont-Tuset, J., Leal-Taix\u00e9, L., Cremers, D., and Van Gool, L. (2017, January 21\u201326). One-Shot Video Object Segmentation. Proceedings of the 30th IEEE Conference on Computer Vision and Pattern Recognition, CVPR 2017, Honolulu, HI, USA.","DOI":"10.1109\/CVPR.2017.565"},{"key":"ref_13","doi-asserted-by":"crossref","unstructured":"Perazzi, F., Pont-Tuset, J., McWilliams, B., Van Gool, L., Gross, M., and Sorkine-Hornung, A. (2016, January 30). A Benchmark Dataset and Evaluation Methodology for Video Object Segmentation. Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition, Las Vegas, NV, USA.","DOI":"10.1109\/CVPR.2016.85"},{"key":"ref_14","doi-asserted-by":"crossref","unstructured":"Perazzi, F., Khoreva, A., Benenson, R., Schiele, B., and Sorkine-Hornung, A. (2017, January 21\u201326). Learning Video Object Segmentation from Static Images. Proceedings of the IEEE Computer Society Conference on Computer Vision and Pattern Recognition, Honolulu, HI, USA.","DOI":"10.1109\/CVPR.2017.372"},{"key":"ref_15","doi-asserted-by":"crossref","unstructured":"Cheng, H.K., Tai, Y.W., and Tang, C.K. (2021, January 20\u201325). Modular Interactive Video Object Segmentation: Interaction-to-Mask, Propagation and Difference-Aware Fusion. Proceedings of the IEEE Computer Society Conference on Computer Vision and Pattern Recognition, Nashville, TN, USA.","DOI":"10.1109\/CVPR46437.2021.00551"},{"key":"ref_16","doi-asserted-by":"crossref","unstructured":"Cheng, H.K., and Schwing, A.G. (2022). XMem: Long-Term Video Object Segmentation with an Atkinson-Shiffrin Memory Model, Springer.","DOI":"10.1007\/978-3-031-19815-1_37"},{"key":"ref_17","doi-asserted-by":"crossref","unstructured":"Li, M., Hu, L., Xiong, Z., Zhang, B., Pan, P., and Liu, D. (2022, January 24). Recurrent Dynamic Embedding for Video Object Segmentation. Proceedings of the IEEE Computer Society Conference on Computer Vision and Pattern Recognition, New Orleans, LA, USA.","DOI":"10.1109\/CVPR52688.2022.00139"},{"key":"ref_18","unstructured":"Liang, Y., Li, X., Jafari, N., and Chen, Q. (2020, January 15). Video Object Segmentation with Adaptive Feature Bank and Uncertain-Region Refinement. Proceedings of the Advances in Neural Information Processing Systems, Vancouver, BC, Canada."},{"key":"ref_19","doi-asserted-by":"crossref","unstructured":"Li, X., and Loy, C.C. (2018). Video Object Segmentation with Joint Re-Identification and Attention-Aware Mask Propagation, Springer.","DOI":"10.1007\/978-3-030-01219-9_6"},{"key":"ref_20","doi-asserted-by":"crossref","unstructured":"Rahmatulloh, A., Gunawan, R., Sulastri, H., Pratama, I., and Darmawan, I. (2021, January 13\u201314). Face Mask Detection Using Haar Cascade Classifier Algorithm Based on Internet of Things with Telegram Bot Notification. Proceedings of the 2021 International Conference Advancement in Data Science, E-Learning and Information Systems, ICADEIS 2021, Nusa Dua Bali, Indonesia.","DOI":"10.1109\/ICADEIS52521.2021.9702065"},{"key":"ref_21","doi-asserted-by":"crossref","first-page":"5667012","DOI":"10.1155\/2022\/5667012","article-title":"SFDWA: Secure and Fault-Tolerant Aware Delay Optimal Workload Assignment Schemes in Edge Computing for Internet of Drone Things Applications","volume":"2022","author":"Lakhan","year":"2022","journal-title":"Wirel. Commun. Mob. Comput."},{"key":"ref_22","doi-asserted-by":"crossref","first-page":"391","DOI":"10.1007\/s00500-021-05613-8","article-title":"An Agent Architecture for Autonomous UAV Flight Control in Object Classification and Recognition Missions","volume":"27","author":"Mostafa","year":"2023","journal-title":"Soft Comput."},{"key":"ref_23","unstructured":"Dosovitskiy, A., Beyer, L., Kolesnikov, A., Weissenborn, D., Zhai, X., Unterthiner, T., Dehghani, M., Minderer, M., Heigold, G., and Gelly, S. (2021, January 3\u20137). An Image Is Worth 16 \u00d7 16 Words: Transformers for Image Recognition At Scale. Proceedings of the ICLR 2021\u20149th International Conference on Learning Representations, Virtual Event, Austria."},{"key":"ref_24","unstructured":"Zhang, C., Han, D., Qiao, Y., Kim, J.U., Bae, S.-H., Lee, S., and Hong, C.S. (2023). Faster Segment Anything: Towards Lightweight SAM for Mobile Applications. arXiv."},{"key":"ref_25","doi-asserted-by":"crossref","unstructured":"Wu, K., Zhang, J., Peng, H., Liu, M., Xiao, B., Fu, J., and Yuan, L. (2022). TinyViT: Fast Pretraining Distillation for Small Vision Transformers, Springer.","DOI":"10.1007\/978-3-031-19803-8_5"},{"key":"ref_26","doi-asserted-by":"crossref","first-page":"640","DOI":"10.1109\/TPAMI.2016.2572683","article-title":"Fully Convolutional Networks for Semantic Segmentation","volume":"39","author":"Shelhamer","year":"2017","journal-title":"IEEE Trans. Pattern Anal. Mach. Intell."},{"key":"ref_27","doi-asserted-by":"crossref","unstructured":"Held, D., Thrun, S., and Savarese, S. (2016). Learning to Track at 100 FPS with Deep Regression Networks, Springer.","DOI":"10.1007\/978-3-319-46448-0_45"},{"key":"ref_28","doi-asserted-by":"crossref","unstructured":"Bolme, D.S., Beveridge, J.R., Draper, B.A., and Lui, Y.M. (2010, January 13\u201318). Visual Object Tracking Using Adaptive Correlation Filters. Proceedings of the IEEE Computer Society Conference on Computer Vision and Pattern Recognition, San Francisco, CA, USA.","DOI":"10.1109\/CVPR.2010.5539960"},{"key":"ref_29","doi-asserted-by":"crossref","first-page":"583","DOI":"10.1109\/TPAMI.2014.2345390","article-title":"High-Speed Tracking with Kernelized Correlation Filters","volume":"37","author":"Henriques","year":"2015","journal-title":"IEEE Trans. Pattern Anal. Mach. Intell."},{"key":"ref_30","doi-asserted-by":"crossref","first-page":"671","DOI":"10.1007\/s11263-017-1061-3","article-title":"Discriminative Correlation Filter Tracker with Channel and Spatial Reliability","volume":"126","author":"Matas","year":"2018","journal-title":"Int. J. Comput. Vis."},{"key":"ref_31","doi-asserted-by":"crossref","first-page":"6551","DOI":"10.3934\/mbe.2023282","article-title":"Deep Learning-Based Small Object Detection: A Survey","volume":"20","author":"Feng","year":"2023","journal-title":"Math. Biosci. Eng."},{"key":"ref_32","unstructured":"Vaswani, A., Shazeer, N., Parmar, N., Uszkoreit, J., Jones, L., Gomez, A.N., Kaiser, \u0141., and Polosukhin, I. (2017, January 4\u20139). Attention Is All You Need. Proceedings of the Advances in Neural Information Processing Systems, Long Beach, CA, USA."},{"key":"ref_33","doi-asserted-by":"crossref","unstructured":"Li, R.Y.M., Tang, B., and Chau, K.W. (2019). Sustainable Construction Safety Knowledge Sharing: A Partial Least Square-Structural Equation Modeling and a Feedforward Neural Network Approach. Sustainability, 11.","DOI":"10.3390\/su11205831"},{"key":"ref_34","doi-asserted-by":"crossref","unstructured":"Nguyen, A., Pham, K., Ngo, D., Ngo, T., and Pham, L. (2021, January 26\u201328). An Analysis of State-of-the-Art Activation Functions for Supervised Deep Neural Network. Proceedings of the 2021 International Conference on System Science and Engineering, ICSSE 2021, Ho Chi Minh City, Vietnam.","DOI":"10.1109\/ICSSE52999.2021.9538437"},{"key":"ref_35","unstructured":"Tancik, M., Srinivasan, P.P., Mildenhall, B., Fridovich-Keil, S., Raghavan, N., Singhal, U., Ramamoorthi, R., Barron, J.T., and Ng, R. (2020, January 6\u201312). Fourier Features Let Networks Learn High Frequency Functions in Low Dimensional Domains. Proceedings of the Advances in Neural Information Processing Systems, Online."},{"key":"ref_36","unstructured":"Dalal, N., Triggs, B., Dalal, N., and Triggs, B. (2005, January 20\u201325). Histograms of Oriented Gradients for Human Detection To Cite This Version: Histograms of Oriented Gradients for Human Detection. Proceedings of the IEEE Computer Society Conference on Computer Vision and Pattern Recognition, San Diego, CA, USA."},{"key":"ref_37","unstructured":"Zamir, S.W., Arora, A., Gupta, A., Khan, S., Sun, G., Khan, F.S., Zhu, F., Shao, L., Xia, G.-S., and Bai, X. (2020, January 14\u201319). ISAID: A Large-Scale Dataset for Instance Segmentation in Aerial Images. Proceedings of the IEEE Computer Society Conference on Computer Vision and Pattern Recognition, Seattle, WA, USA."},{"key":"ref_38","doi-asserted-by":"crossref","unstructured":"Xia, G.S., Bai, X., Ding, J., Zhu, Z., Belongie, S., Luo, J., Datcu, M., Pelillo, M., and Zhang, L. (2018, January 18\u201323). DOTA: A Large-Scale Dataset for Object Detection in Aerial Images. Proceedings of the IEEE Computer Society Conference on Computer Vision and Pattern Recognition, Salt Lake City, UT, USA.","DOI":"10.1109\/CVPR.2018.00418"},{"key":"ref_39","doi-asserted-by":"crossref","unstructured":"Shermeyer, J., Hossler, T., Van Etten, A., Hogan, D., Lewis, R., and Kim, D. (2021, January 3\u20138). RarePlanes: Synthetic Data Takes Flight. Proceedings of the 2021 IEEE Winter Conference on Applications of Computer Vision (WACV), Waikoloa, HI, USA.","DOI":"10.1109\/WACV48630.2021.00025"},{"key":"ref_40","doi-asserted-by":"crossref","unstructured":"Li, F., Kim, T., Humayun, A., Tsai, D., and Rehg, J.M. (2013, January 1\u20138). Video Segmentation by Tracking Many Figure-Ground Segments. Proceedings of the IEEE International Conference on Computer Vision, Sydney, Australia.","DOI":"10.1109\/ICCV.2013.273"},{"key":"ref_41","unstructured":"Pont-Tuset, J., Perazzi, F., Caelles, S., Arbel\u00e1ez, P., Sorkine-Hornung, A., and Van Gool, L. (2017). The 2017 DAVIS Challenge on Video Object Segmentation. arXiv."},{"key":"ref_42","first-page":"603","article-title":"YouTube-VOS: Sequence-to-Sequence Video Object Segmentation","volume":"Volume 3","author":"Xu","year":"2018","journal-title":"Scanning Microscopy"},{"key":"ref_43","doi-asserted-by":"crossref","unstructured":"Oh, S.W., Lee, J.Y., Sunkavalli, K., and Kim, S.J. (2018, January 18\u201323). Fast Video Object Segmentation by Reference-Guided Mask Propagation. Proceedings of the IEEE Computer Society Conference on Computer Vision and Pattern Recognition, Salt Lake City, UT, USA.","DOI":"10.1109\/CVPR.2018.00770"},{"key":"ref_44","doi-asserted-by":"crossref","first-page":"149","DOI":"10.1016\/j.asoc.2017.07.053","article-title":"Bio-Inspired Computation: Recent Development on the Modifications of the Cuckoo Search Algorithm","volume":"61","author":"Chiroma","year":"2017","journal-title":"Appl. Soft Comput. J."},{"key":"ref_45","first-page":"8507","article-title":"High-Performance Transformer Tracking","volume":"45","author":"Chen","year":"2023","journal-title":"IEEE Trans. Pattern Anal. Mach. Intell."},{"key":"ref_46","first-page":"6168","article-title":"Robust Online Tracking with Meta-Updater","volume":"45","author":"Zhao","year":"2023","journal-title":"IEEE Trans. Pattern Anal. Mach. Intell."},{"key":"ref_47","doi-asserted-by":"crossref","unstructured":"Zhu, J., Lai, S., Chen, X., Wang, D., and Lu, H. (2023, January 17\u201324). Visual Prompt Multi-Modal Tracking. Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition, Vancouver, BC, Canada.","DOI":"10.1109\/CVPR52729.2023.00918"},{"key":"ref_48","doi-asserted-by":"crossref","unstructured":"Chen, X., Peng, H., Wang, D., Lu, H., and Hu, H. (2023, January 18\u201322). SeqTrack: Sequence to Sequence Learning for Visual Object Tracking. Proceedings of the IEEE Computer Society Conference on Computer Vision and Pattern Recognition, Vancouver, BC, Canada.","DOI":"10.1109\/CVPR52729.2023.01400"},{"key":"ref_49","doi-asserted-by":"crossref","unstructured":"Liu, S., Li, X., Lu, H., and He, Y. (2022, January 18\u201324). Multi-Object Tracking Meets Moving UAV. Proceedings of the IEEE Computer Society Conference on Computer Vision and Pattern Recognition, New Orleans, LA, USA.","DOI":"10.1109\/CVPR52688.2022.00867"},{"key":"ref_50","doi-asserted-by":"crossref","unstructured":"Li, R., He, C., Li, S., Zhang, Y., and Zhang, L. (2023, January 18\u201322). DynaMask: Dynamic Mask Selection for Instance Segmentation. Proceedings of the IEEE Computer Society Conference on Computer Vision and Pattern Recognition, Vancouver, BC, Canada.","DOI":"10.1109\/CVPR52729.2023.01085"},{"key":"ref_51","doi-asserted-by":"crossref","unstructured":"Li, R., He, C., Zhang, Y., Li, S., Chen, L., and Zhang, L. (2023, January 18\u201322). SIM: Semantic-Aware Instance Mask Generation for Box-Supervised Instance Segmentation. Proceedings of the IEEE Computer Society Conference on Computer Vision and Pattern Recognition, Vancouver, BC, Canada.","DOI":"10.1109\/CVPR52729.2023.00695"},{"key":"ref_52","doi-asserted-by":"crossref","unstructured":"Zhang, T., Wei, S., and Ji, S. (2022, January 18\u201324). E2EC: An End-to-End Contour-Based Method for High-Quality High-Speed Instance Segmentation. Proceedings of the IEEE Computer Society Conference on Computer Vision and Pattern Recognition, New Orleans, LA, USA.","DOI":"10.1109\/CVPR52688.2022.00440"},{"key":"ref_53","doi-asserted-by":"crossref","unstructured":"Zhu, C., Zhang, X., Li, Y., Qiu, L., Han, K., and Han, X. (2022, January 18\u201324). SharpContour: A Contour-Based Boundary Refinement Approach for Efficient and Accurate Instance Segmentation. Proceedings of the IEEE Computer Society Conference on Computer Vision and Pattern Recognition, New Orleans, LA, USA.","DOI":"10.1109\/CVPR52688.2022.00435"},{"key":"ref_54","doi-asserted-by":"crossref","unstructured":"Cheng, T., Wang, X., Chen, S., Zhang, W., Zhang, Q., Huang, C., Zhang, Z., and Liu, W. (2022, January 18\u201324). Sparse Instance Activation for Real-Time Instance Segmentation. Proceedings of the IEEE Computer Society Conference on Computer Vision and Pattern Recognition, New Orleans, LA, USA.","DOI":"10.1109\/CVPR52688.2022.00439"}],"container-title":["Remote Sensing"],"original-title":[],"language":"en","link":[{"URL":"https:\/\/www.mdpi.com\/2072-4292\/15\/24\/5665\/pdf","content-type":"unspecified","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2025,10,10]],"date-time":"2025-10-10T21:35:00Z","timestamp":1760132100000},"score":1,"resource":{"primary":{"URL":"https:\/\/www.mdpi.com\/2072-4292\/15\/24\/5665"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2023,12,7]]},"references-count":54,"journal-issue":{"issue":"24","published-online":{"date-parts":[[2023,12]]}},"alternative-id":["rs15245665"],"URL":"https:\/\/doi.org\/10.3390\/rs15245665","relation":{},"ISSN":["2072-4292"],"issn-type":[{"value":"2072-4292","type":"electronic"}],"subject":[],"published":{"date-parts":[[2023,12,7]]}}}