{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2026,7,7]],"date-time":"2026-07-07T15:55:58Z","timestamp":1783439758430,"version":"3.54.6"},"reference-count":75,"publisher":"Springer Science and Business Media LLC","issue":"4","license":[{"start":{"date-parts":[[2024,4,12]],"date-time":"2024-04-12T00:00:00Z","timestamp":1712880000000},"content-version":"tdm","delay-in-days":0,"URL":"https:\/\/creativecommons.org\/licenses\/by\/4.0"},{"start":{"date-parts":[[2024,4,12]],"date-time":"2024-04-12T00:00:00Z","timestamp":1712880000000},"content-version":"vor","delay-in-days":0,"URL":"https:\/\/creativecommons.org\/licenses\/by\/4.0"}],"content-domain":{"domain":[],"crossmark-restriction":false},"short-container-title":["Mach. Intell. Res."],"published-print":{"date-parts":[[2024,8]]},"abstract":"<jats:title>Abstract<\/jats:title>\n                  <jats:p>\n                    Recently, Meta AI Research approaches a general, promptable segment anything model (SAM) pre-trained on an unprecedentedly large segmentation dataset (SA-1B). Without a doubt, the emergence of SAM will yield significant benefits for a wide array of practical image segmentation applications. In this study, we conduct a series of intriguing investigations into the performance of SAM across various applications, particularly in the fields of natural images, agriculture, manufacturing, remote sensing and healthcare. We analyze and discuss the benefits and limitations of SAM, while also presenting an outlook on its future development in segmentation tasks. By doing so, we aim to give a comprehensive understanding of SAM\u2019s practical applications. This work is expected to provide insights that facilitate future research activities toward generic segmentation. Source code is publicly available at\n                    <jats:ext-link xmlns:xlink=\"http:\/\/www.w3.org\/1999\/xlink\" xlink:href=\"https:\/\/github.com\/LiuTingWed\/SAM-Not-Perfect\" ext-link-type=\"uri\">https:\/\/github.com\/LiuTingWed\/SAM-Not-Perfect<\/jats:ext-link>\n                    .\n                  <\/jats:p>","DOI":"10.1007\/s11633-023-1385-0","type":"journal-article","created":{"date-parts":[[2024,4,12]],"date-time":"2024-04-12T01:01:31Z","timestamp":1712883691000},"page":"617-630","source":"Crossref","is-referenced-by-count":130,"title":["Segment Anything Is Not Always Perfect: An Investigation of SAM on Different Real-world Applications"],"prefix":"10.1007","volume":"21","author":[{"ORCID":"https:\/\/orcid.org\/0000-0003-4059-5902","authenticated-orcid":false,"given":"Wei","family":"Ji","sequence":"first","affiliation":[],"role":[{"vocabulary":"crossref","role":"author"}]},{"ORCID":"https:\/\/orcid.org\/0000-0003-0811-4988","authenticated-orcid":false,"given":"Jingjing","family":"Li","sequence":"additional","affiliation":[],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Qi","family":"Bi","sequence":"additional","affiliation":[],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Tingwei","family":"Liu","sequence":"additional","affiliation":[],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Wenbo","family":"Li","sequence":"additional","affiliation":[],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Li","family":"Cheng","sequence":"additional","affiliation":[],"role":[{"vocabulary":"crossref","role":"author"}]}],"member":"297","published-online":{"date-parts":[[2024,4,12]]},"reference":[{"key":"1385_CR1","unstructured":"A. Radford, J. W. Kim, C. Hallacy, A. Ramesh, G. Goh, S. Agarwal, G. Sastry, A. Askell, P. Mishkin, J. Clark, G. Krueger, I. Sutskever. Learning transferable visual models from natural language supervision. In Proceedings of the 38th International Conference on Machine Learning, pp. 8748\u20138763, 2021."},{"key":"1385_CR2","doi-asserted-by":"crossref","unstructured":"A. Kirillov, E. Mintun, N. Ravi, H. Mao, C. Rolland, L. Gustafson, T. Xiao, S. Whitehead, A. C. Berg, W. Y. Lo, P. Dollar, R. Girshick. Segment anything, [Online], Available: https:\/\/arxiv.org\/abs\/2304.02643, 2023.","DOI":"10.1109\/ICCV51070.2023.00371"},{"issue":"6","key":"1385_CR3","doi-asserted-by":"publisher","first-page":"1737","DOI":"10.3390\/s20061737","volume":"20","author":"T Y Ko","year":"2020","unstructured":"T. Y. Ko, S. H. Lee. Novel method of semantic segmentation applicable to augmented reality. Sensors, vol. 20, no. 6, pp. 1737, 2020. DOI: https:\/\/doi.org\/10.3390\/s20061737.","journal-title":"Sensors"},{"key":"1385_CR4","unstructured":"B. Wang, A. Aboah, Z. Y. Zhang, U. Bagci. GazeSAM: What you see is what you segment, [Online], Available: https:\/\/arxiv.org\/abs\/2304_13844, 2023"},{"issue":"2","key":"1385_CR5","doi-asserted-by":"publisher","first-page":"117","DOI":"10.1007\/s41095-019-0149-9","volume":"5","author":"A Borji","year":"2019","unstructured":"A. Borji, M. M. Cheng, Q. B. Hou, H. Z. Jiang, J. Li. Salient object detection: A survey. Computational Visual Media, vol. 5, no. 2, pp. 117\u2013150, 2019. DOI: https:\/\/doi.org\/10.1007\/s41095-019-0149-9.","journal-title":"Computational Visual Media"},{"key":"1385_CR6","doi-asserted-by":"publisher","unstructured":"W. Ji, S. Yu, J. D. Wu, K. Ma, C. Bian, Q. Bi, J. J. Li, H. R. Liu, L. Cheng, Y. F. Zheng. Learning calibrated medical image segmentation via multi-rater agreement modeling. In Proceedings of the IEEE\/CVF Conference on Computer Vision and Pattern Recognition, Nashvllle, USA, pp. 12336\u201312346, 2021. DOI: https:\/\/doi.org\/10.1109\/CVPR46437.2021.01216.","DOI":"10.1109\/CVPR46437.2021.01216"},{"key":"1385_CR7","doi-asserted-by":"publisher","first-page":"123453","DOI":"10.1109\/ACCESS.2019.2937461","volume":"7","author":"T He","year":"2019","unstructured":"T. He, Y. Liu, C. Y. Xu, X. L. Zhou, Z. K. Hu, J. N. Fan. A fully convolutional neural network for wood defect location and identification. IEEE Access, vol. 7, pp. 123453\u2013123462, 2019. DOI: https:\/\/doi.org\/10.1109\/ACCESS.2019.2937461.","journal-title":"IEEE Access"},{"key":"1385_CR8","doi-asserted-by":"publisher","unstructured":"Y. N. Li, Z. Y. Huang, Z. G. Cao, H. Lu, H. H. Wang, S. P. Zhang. Performance evaluation of crop segmentation algorithms. IEEE Access, vol.8, pp.36210\u201336225, 2020. DOI: https:\/\/doi.org\/10.1109\/ACCESS.2020.2969451.","DOI":"10.1109\/ACCESS.2020.2969451"},{"key":"1385_CR9","doi-asserted-by":"publisher","unstructured":"Y. Y. Xu, Z. Xie, Y. X. Feng, Z. L. Chen. Road extraction from high-resolution remote sensing imagery using deep learning. Remote Sensing, vol. 10, no. 9, Article number 1461, 2018. DOI: https:\/\/doi.org\/10.3390\/rs10091461.","DOI":"10.3390\/rs10091461"},{"key":"1385_CR10","unstructured":"J. Li, W. Ji, S. Wang, W. Li, L. Cheng. DVSOD: RGB-D video salient object detection. In Proceedings of the Advances in Neural Information Processing Systems, New Orleans, USA, 2023."},{"issue":"12","key":"1385_CR11","doi-asserted-by":"publisher","first-page":"9026","DOI":"10.1109\/TPAMI.2021.3122139","volume":"44","author":"N Liu","year":"2022","unstructured":"N. Liu, N. Zhang, L. Shao, J. W. Han. Learning selective mutual attention and contrast for RGB-D saliency detection. IEEE Transactions on Pattern Analysis and Machine Intelligence, vol. 44, no. 12, pp. 9026\u20139042, 2022. DOI: https:\/\/doi.org\/10.1109\/TPAMI.2021.3122139.","journal-title":"IEEE Transactions on Pattern Analysis and Machine Intelligence"},{"key":"1385_CR12","doi-asserted-by":"publisher","unstructured":"D. P. Fan, G. P. Ji, G. L. Sun, M. M. Cheng, J. B. Shen, L. Shao. Camouflaged object detection. In Proceedings of the IEEE\/CVF Conference on Computer Vision and Pattern Recognition, Seattle, USA, pp.2774\u20132784, 2020. DOI: https:\/\/doi.org\/10.1109\/CVPR42600.2020.00285.","DOI":"10.1109\/CVPR42600.2020.00285"},{"issue":"12","key":"1385_CR13","doi-asserted-by":"publisher","first-page":"226101","DOI":"10.1007\/s11432-023-3881-x","volume":"66","author":"G P Ji","year":"2023","unstructured":"G. P. Ji, D. P. Fan, P. Xu, B. W. Zhou, M. M. Cheng, L. Van Gool. SAM struggles in concealed scenes-empirical study on \u201csegment anything\u201d. Science China Information Sciences, vol. 66, no. 12, pp.226101, 2023. DOI: https:\/\/doi.org\/10.1007\/s11432-023-3881-x.","journal-title":"Science China Information Sciences"},{"key":"1385_CR14","doi-asserted-by":"publisher","unstructured":"E. Z. Xie, W. J. Wang, W. H. Wang, P. Z. Sun, H. Xu, D. Liang, P. Luo. Segmenting transparent objects in the wild with transformer. In Proceedings of the 30th International Joint Conference on Artificial Intelligence, pp. 1194\u20131200, 2021. DOI: https:\/\/doi.org\/10.24963\/ijcai.2021\/165.","DOI":"10.24963\/ijcai.2021\/165"},{"key":"1385_CR15","doi-asserted-by":"publisher","first-page":"1925","DOI":"10.1109\/TIP.2021.3049331","volume":"30","author":"X W Hu","year":"2021","unstructured":"X. W. Hu, T. Y. Wang, C. W. Fu, Y. T. Jiang, Q. Wang, P. A. Heng. Revisiting shadow detection: A new benchmark dataset for complex world. IEEE Transactions on Image Processing, vol. 30, pp. 1925\u20131934, 2021. DOI: https:\/\/doi.org\/10.1109\/TIP.2021.3049331.","journal-title":"IEEE Transactions on Image Processing"},{"issue":"4","key":"1385_CR16","doi-asserted-by":"publisher","first-page":"1337","DOI":"10.1109\/TPAMI.2019.2948011","volume":"43","author":"L Hou","year":"2021","unstructured":"L. Hou, T. F. Y. Vicente, M. Hoai, D. Samaras. Large scale shadow annotation and detection using lazy annotation and stacked CNNS. IEEE Transactions on Pattern Analysis and Machine Intelligence, vol. 43, no. 4, pp.1337\u20131351, 2021. DOI: https:\/\/doi.org\/10.1109\/TPAMI.2019.2948011.","journal-title":"IEEE Transactions on Pattern Analysis and Machine Intelligence"},{"key":"1385_CR17","doi-asserted-by":"publisher","first-page":"58","DOI":"10.1016\/j.compag.2013.04.010","volume":"96","author":"W Guo","year":"2013","unstructured":"W. Guo, U. K. Rage, S. Ninomiya. Illumination invariant segmentation of vegetation for time series wheat images based on decision tree model. Computers and Electronics in Agriculture, vol. 96, pp.58\u201366, 2013. DOI: https:\/\/doi.org\/10.1016\/j.compag.2013.04.010.","journal-title":"Computers and Electronics in Agriculture"},{"key":"1385_CR18","doi-asserted-by":"publisher","unstructured":"A. Sriwastwa, S. Prakash, S. Swarit, K. Kumari, S. S. Sahu. Detection of pests using color based image segmenttion. In Proceedings of the 2nd International Conference on Inventive Communication and Computational Technologies, Coimbatore, India, pp. 1393\u20131396, 2018. DOI: https:\/\/doi.org\/10.1109\/ICICCT.2018.8473166.","DOI":"10.1109\/ICICCT.2018.8473166"},{"key":"1385_CR19","unstructured":"D. Contributors. Leaf disease segmentation dataset, [Online], Available: https:\/\/www.kaggle.com\/datasets\/fakh-realam9537\/leaf-disease-segmentation-dataset, 2023."},{"key":"1385_CR20","doi-asserted-by":"publisher","unstructured":"P. Bergmann, M. Fauser, D. Sattlegger, C. Steger. MVTec AD-A comprehensive real-world dataset for unsupervised anomaly detection. In Proceedings of the IEEE\/CVF Conference on Computer Vision and Pattern Recognition, Long Beach, USA, pp. 9584\u20139592, 2019. DOI: https:\/\/doi.org\/10.1109\/CV-PR.2019.00982.","DOI":"10.1109\/CV-PR.2019.00982"},{"issue":"4","key":"1385_CR21","doi-asserted-by":"publisher","first-page":"760","DOI":"10.3390\/rs13040760","volume":"13","author":"S He","year":"2021","unstructured":"S. He, W. S. Jiang. Boundary-assisted learning for building extraction from optical remote sensing imagery. Remote Sensing, vol. 13, no.4, pp.760, 2021. DOI: https:\/\/doi.org\/10.3390\/rs13040760.","journal-title":"Remote Sensing"},{"key":"1385_CR22","doi-asserted-by":"publisher","first-page":"6498","DOI":"10.1109\/TIP.2021.3092816","volume":"30","author":"Q Bi","year":"2021","unstructured":"Q. Bi, K. Qin, H. Zhang, G. S. Xia. Local semantic enhanced convNet for aerial scene recognition. IEEE Transactions on Image Processing, vol. 30, pp.6498\u20136511, 2021. DOI: https:\/\/doi.org\/10.1109\/TIP.2021.3092816.","journal-title":"IEEE Transactions on Image Processing"},{"key":"1385_CR23","first-page":"1","volume-title":"Machine Learning for Aerial Image Labeling","author":"V Mnih","year":"2013","unstructured":"V. Mnih, G. Hinton. Machine Learning for Aerial Image Labeling, Toronto, Canada: University of Toronto, pp. 1\u201324, 2013."},{"issue":"7","key":"1385_CR24","doi-asserted-by":"publisher","first-page":"1597","DOI":"10.1109\/TMI.2018.2791488","volume":"37","author":"H Z Fu","year":"2018","unstructured":"H. Z. Fu, J. Cheng, Y. W. Xu, D. W. K. Wong, J. Liu, X. C. Cao. Joint optic disc and cup segmentation based on multi-label deep network and polar transformation. IEEE Transactions on Medical Imaging, vol. 37, no.7, pp. 1597\u20131605, 2018. DOI: https:\/\/doi.org\/10.1109\/TMI.2018.2791488.","journal-title":"IEEE Transactions on Medical Imaging"},{"issue":"3","key":"1385_CR25","doi-asserted-by":"publisher","first-page":"701","DOI":"10.1007\/s10792-016-0329-x","volume":"37","author":"A Almazroa","year":"2017","unstructured":"A. Almazroa, S. Alodhayb, E. Osman, E. Ramadan, M. Hummadi, M. Dlaim, M. Alkatee, K. Raahemifar, V. Lakshminarayanan. Agreement among ophthalmologists in marking the optic disc and optic cup in fundus images. International Ophthalmology, vol. 37, no. 3, pp. 701\u2013717, 2017. DOI: https:\/\/doi.org\/10.1007\/s10792-016-0329-x.","journal-title":"International Ophthalmology"},{"key":"1385_CR26","doi-asserted-by":"publisher","unstructured":"D. P. G. Fan P. Ji, T. Zhou, G. Chen, H. Fu, J. Shen, L. Shao. PraNet: Parallel reverse attention network for polyp segmentation. In Proceedings of the 23rd International Conference on Medical Image Computing and Computer Assisted Intervention, pp. 263\u2013273, 2020. DOI: https:\/\/doi.org\/10.1007\/978-3-030-59725-2_26.","DOI":"10.1007\/978-3-030-59725-2_26"},{"issue":"6","key":"1385_CR27","doi-asserted-by":"publisher","first-page":"531","DOI":"10.1007\/s11633-022-1371-y","volume":"19","author":"G P Ji","year":"2022","unstructured":"G. P. Ji, G. B. Xiao, Y. C. Chou, D. P. Fan, K. Zhao, G. Chen, L. Van Gool. Video polyp segmentation: A deep learning perspective. Machine Intelligence Research, vol. 19, no. 6, pp. 531\u2013549, 2022. DOI: https:\/\/doi.org\/10.1007\/s11633-022-1371-y.","journal-title":"Machine Intelligence Research"},{"key":"1385_CR28","doi-asserted-by":"publisher","unstructured":"J. W. Pan, Q. Bi, Y. Z. Yang, P. F. Zhu, C. Bian. Label-efficient hybrid-supervised learning for medical image segmentation. In Proceedings of the AAAI Conference on Artificial Intelligence, pp. 2026\u20132034, 2022. DOI: https:\/\/doi.org\/10.1609\/aaai.v36i2.20098.","DOI":"10.1609\/aaai.v36i2.20098"},{"key":"1385_CR29","doi-asserted-by":"publisher","first-page":"43669","DOI":"10.1109\/ACCESS.2022.3168693","volume":"10","author":"N S An","year":"2022","unstructured":"N. S. An, P. N. Lan, D. V. Hang, D. V. Long, T. Q. Trung, N. T. Thuy, D. V. Sang. Blazeneo: Blazing fast polyp segmentation and neoplasm detection. IEEE Access, vol. 10, pp. 43669\u201343684, 2022. DOI: https:\/\/doi.org\/10.1109\/ACCESS.2022.3168693.","journal-title":"IEEE Access"},{"key":"1385_CR30","doi-asserted-by":"publisher","unstructured":"N. C. F. Codella, D. Gutman, M. E. Celebi, B. Helba, M. A. Marchetti, S. W. Dusza, A. Kalloo, K. Liopyris, N. Mishra, H. Kittler, A. Halpern. Skin lesion analysis toward melanoma detection: A challenge at the 2017 international symposium on biomedical imaging (ISBI), hosted by the international skin imaging collaboration (ISIC). In Proceedings of the 15th International Symposium on Biomedical Imaging, Washington DC, USA, pp.168\u2013172, 2018. DOI: https:\/\/doi.org\/10.1109\/ISBI.2018.8363547.","DOI":"10.1109\/ISBI.2018.8363547"},{"key":"1385_CR31","doi-asserted-by":"publisher","unstructured":"T. Mendon\u00e7a, P. M. Ferreira, J. S. Marques, A. R. S. Marcal, J. Rozeira. PH2- a dermoscopic image database for research and benchmarking. In Proceedings of the 35th Annual International Conference of the IEEE Engineering in Medicine and Biology Society, Osaka, Japan, pp. 5437\u20135440, 2013. DOI: https:\/\/doi.org\/10.1109\/EMBC.2013.6610779.","DOI":"10.1109\/EMBC.2013.6610779"},{"key":"1385_CR32","doi-asserted-by":"publisher","unstructured":"L. J. Wang, H. C. Lu, Y. F. Wang, M. Y. Feng, D. Wang, B. C. Yin, X. Ruan. Learning to detect salient objects with image-level supervision. In Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition, Honolulu, USA, pp.3796\u20133805, 2017. DOI: https:\/\/doi.org\/10.1109\/CVPR.2017.404.","DOI":"10.1109\/CVPR.2017.404"},{"key":"1385_CR33","doi-asserted-by":"publisher","unstructured":"J. Zhang, D. P. Fan, Y. C. Dai, X. Yu, Y. R. Zhong, N. Barnes, L. Shao. RGB-D saliency detection via cascaded mutual information minimization. In Proceedings of the IEEE\/CVF International Conference on Computer Vision, Montreal, Canada, pp. 4318\u20134327, 2021. DOI: https:\/\/doi.org\/10.1109\/IC-CV48922.2021.00430.","DOI":"10.1109\/IC-CV48922.2021.00430"},{"issue":"1","key":"1385_CR34","doi-asserted-by":"publisher","first-page":"160","DOI":"10.1109\/TMM.2019.2924578","volume":"22","author":"Z Z Tu","year":"2020","unstructured":"Z. Z. Tu, T. Xia, C. L. Li, X. X. Wang, Y. Ma, J. Tang. RGB-T image saliency detection via collaborative graph learning. IEEE Transactions on Multimedia, vol. 22, no. 1, pp. 160\u2013173, 2020. DOI: https:\/\/doi.org\/10.1109\/TMM.2019.2924578.","journal-title":"IEEE Transactions on Multimedia"},{"key":"1385_CR35","doi-asserted-by":"publisher","unstructured":"X. B. Qin, H. Dai, X. B. Hu, D. P. Fan, L. Shao, L. Van Gool. Highly accurate dichotomous image segmentation. In Proceedings of the 17th European Conference on Computer Vision, Tel Aviv, Israel, pp.38\u201356, 2022 DOI: https:\/\/doi.org\/10.1007\/978-3-031-19797-0_3.","DOI":"10.1007\/978-3-031-19797-0_3"},{"key":"1385_CR36","doi-asserted-by":"publisher","unstructured":"T. F. Y. Vicente, L. Hou, C. P. Yu, M. Hoai, D. Samaras. Large-scale training of shadow detectors with noisily-annotated shadow examples. In Proceedings of the 14th European Conference on Computer Vision, Amsterdam, The Netherlands, pp.816\u2013832, 2016. DOI: https:\/\/doi.org\/10.1007\/978-3-319-46466-4_49.","DOI":"10.1007\/978-3-319-46466-4_49"},{"issue":"1","key":"1385_CR37","doi-asserted-by":"publisher","first-page":"16","DOI":"10.1007\/s44267-023-00019-6","volume":"1","author":"D P Fan","year":"2023","unstructured":"D. P. Fan, G. P. Ji, P. Xu, M. M. Cheng, C. Sakaridis, L. Van Gool. Advances in deep concealed scene understanding. Visual Intelligence, vol. 1, no. 1, pp. 16, 2023. DOI: https:\/\/doi.org\/10.1007\/s44267-023-00019-6.","journal-title":"Visual Intelligence"},{"issue":"2","key":"1385_CR38","doi-asserted-by":"publisher","first-page":"630","DOI":"10.1109\/TMI.2015.2487997","volume":"35","author":"N Tajbakhsh","year":"2016","unstructured":"N. Tajbakhsh, S. R. Gurudu, J. M. Liang. Automated polyp detection in colonoscopy videos using shape and context information. IEEE Transactions on Medical Imaging, vol. 35, no. 2, pp. 630\u2013644, 2016. DOI: https:\/\/doi.org\/10.1109\/TMI.2015.2487997.","journal-title":"IEEE Transactions on Medical Imaging"},{"key":"1385_CR39","doi-asserted-by":"publisher","unstructured":"N. Liu, N. Zhang, K. Y. Wan, L. Shao, J. W. Han. Visual saliency transformer. In Proceedings of the IEEE\/CVF International Conference on Computer Vision, Montreal, Canada, pp. 4702\u20134712, 2021. DOI: https:\/\/doi.org\/10.1109\/ICCV48922.2021.00468.","DOI":"10.1109\/ICCV48922.2021.00468"},{"issue":"3","key":"1385_CR40","doi-asserted-by":"publisher","first-page":"3738","DOI":"10.1109\/TPAMI.2022.3179526","volume":"45","author":"M C Zhuge","year":"2023","unstructured":"M. C. Zhuge, D. P. Fan, N. Liu, D. W. Zhang, D. Xu, L. Shao. Salient object detection via integrity learning. IEEE Transactions on Pattern Analysis and Machine Intelligence, vol. 45, no. 3, pp. 3738\u20133752, 2023. DOI: https:\/\/doi.org\/10.1109\/TPAMI.2022.3179526.","journal-title":"IEEE Transactions on Pattern Analysis and Machine Intelligence"},{"key":"1385_CR41","doi-asserted-by":"publisher","first-page":"3125","DOI":"10.1109\/TIP.2022.3164550","volume":"31","author":"Y H Wu","year":"2022","unstructured":"Y. H. Wu, Y. Liu, L. Zhang, M. M. Cheng, B. Ren. EDN: Salient object detection via extremely-downsampled network. IEEE Transactions on Image Processing, vol. 31, pp. 3125\u20133136, 2022. DOI: https:\/\/doi.org\/10.1109\/TIP.2022.3164550.","journal-title":"IEEE Transactions on Image Processing"},{"key":"1385_CR42","doi-asserted-by":"publisher","unstructured":"X. Q. Zhao, Y. W. Pang, L. H. Zhang, H. C. Lu, L. Zhang. Suppress and balance: A simple gated network for salient object detection. In Proceedings of the 16th European Conference on Computer Vision, Glasgow, UK, pp. 35\u201351, 2020. DOI: https:\/\/doi.org\/10.1007\/978-3-030-58536-5_3.","DOI":"10.1007\/978-3-030-58536-5_3"},{"key":"1385_CR43","doi-asserted-by":"publisher","unstructured":"H. Y. Mei, G. P. Ji, Z. Q. Wei, X. Yang, X. P. Wei, D. P. Fan. Camouflaged object segmentation with distraction mining. In Proceedings of the IEEE\/CVF Conference on Computer Vision and Pattern Recognition, Nashville, USA, pp.8768\u20138777, 2021. DOI: https:\/\/doi.org\/10.1109\/CVPR46437.2021.00866.","DOI":"10.1109\/CVPR46437.2021.00866"},{"key":"1385_CR44","doi-asserted-by":"publisher","unstructured":"Q. Jia, S. L. Yao, Y. Liu, X. Fan, R. S. Liu, Z. X. Luo. Segment, magnify and reiterate: Detecting camouflaged objects the hard way. In Proceedings of the IEEE\/CVF Conference on Computer Vision and Pattern Recognition, New Orleans, USA, pp. 4703\u20134712, 2022. DOI: https:\/\/doi.org\/10.1109\/CVPR52688.2022.00467.","DOI":"10.1109\/CVPR52688.2022.00467"},{"key":"1385_CR45","doi-asserted-by":"crossref","unstructured":"H. Y. Mei, X. Yang, Y. D. Zhou, G. P. Ji, X. P. Wei, D. P. Fan. Distraction-aware camouflaged object segmentation. Scientia Sinica Informations, 2023","DOI":"10.1360\/SSI-2022-0138"},{"key":"1385_CR46","doi-asserted-by":"publisher","unstructured":"Y. W. Pang, X. Q. Zhao, T. Z. Xiang, L. H. Zhang, H. C. Lu. Zoom in and out: A mixed-scale triplet network for camouflaged object detection. In Proceedings of the IEEE\/CVF Conference on Computer Vision and Pattern Recognition, New Orleans, USA, pp. 2150\u20132160, 2022. DOI: https:\/\/doi.org\/10.1109\/CVPR52688.2022.00220.","DOI":"10.1109\/CVPR52688.2022.00220"},{"key":"1385_CR47","doi-asserted-by":"publisher","unstructured":"X. W. Hu, L. Zhu, C. W. Fu, J. Qin, P. A. Heng. Direction-aware spatial context features for shadow detection. In Proceedings of the IEEE\/CVF Conference on Computer Vision and Pattern Recognition, Salt Lake City, USA, pp. 7454\u20137462, 2018. DOI: https:\/\/doi.org\/10.1109\/CVPR.2018.00778.","DOI":"10.1109\/CVPR.2018.00778"},{"key":"1385_CR48","doi-asserted-by":"publisher","unstructured":"Q. L. Zheng, X. T. Qiao, Y. Cao, R. W. H. Lau. Distraction-aware shadow detection. In Proceedings of the IEEE\/CVF Conference on Computer Vision and Pattern Recognition, Long Beach, USA, pp. 5162\u20135171, 2019. DOI: https:\/\/doi.org\/10.1109\/CVPR.2019.00531.","DOI":"10.1109\/CVPR.2019.00531"},{"key":"1385_CR49","doi-asserted-by":"publisher","unstructured":"Z. H. Chen, L. Zhu, L. Wan, S. Wang, W. Feng, P. A. Heng. A multi-task mean teacher for semi-supervised shadow detection. In Proceedings of the IEEE\/CVF Conference on Computer Vision and Pattern Recognition, Seattle, USA, pp. 5610\u20135619, 2020. DOI: https:\/\/doi.org\/10.1109\/CVPR42600.2020.00565.","DOI":"10.1109\/CVPR42600.2020.00565"},{"issue":"10","key":"1385_CR50","doi-asserted-by":"publisher","first-page":"6024","DOI":"10.1109\/TPAMI.2021.3085766","volume":"44","author":"D P Fan","year":"2022","unstructured":"D. P. Fan, G. P. Ji, M. M. Cheng, L. Shao. Concealed object detection. IEEE Transactions on Pattern Analysis and Machine Intelligence, vol. 44, no. 10, pp. 6024\u20136042, 2022. DOI: https:\/\/doi.org\/10.1109\/TPAMI.2021.3085766.","journal-title":"IEEE Transactions on Pattern Analysis and Machine Intelligence"},{"key":"1385_CR51","doi-asserted-by":"publisher","unstructured":"X. B. Hu, S. Wang, X. B. Qin, H. Dai, W. Q. Ren, D. H. Luo, Y. Tai, L. Shao. High-resolution iterative feedback network for camouflaged object detection. In Proceedings of the AAAI Conference on Artificial Intelligence, Washington DC, USA, pp. 881\u2013889, 2023. DOI: https:\/\/doi.org\/10.1609\/aaai.v37i1.25167.","DOI":"10.1609\/aaai.v37i1.25167"},{"issue":"1","key":"1385_CR52","doi-asserted-by":"publisher","first-page":"92","DOI":"10.1007\/s11633-022-1365-9","volume":"20","author":"G P Ji","year":"2023","unstructured":"G. P. Ji, D. P. Fan, Y. C. Chou, D. X. Dai, A. Liniger, L. Van Gool. Deep gradient learning for efficient camouflaged object detection. Machine Intelligence Research, vol. 20, no.1, pp. 92\u2013108, 2023. DOI: https:\/\/doi.org\/10.1007\/s11633-022-1365-9.","journal-title":"Machine Intelligence Research"},{"key":"1385_CR53","doi-asserted-by":"publisher","first-page":"7036","DOI":"10.1109\/TIP.2022.3217695","volume":"31","author":"T Zhou","year":"2022","unstructured":"T. Zhou, Y. Zhou, C. Gong, J. Yang, Y. Zhang. Feature aggregation and propagation network for camouflaged object detection. IEEE Transactions on Image Processing, vol. 31, pp.7036\u20137047, 2022. OOI: https:\/\/doi.org\/10.1109\/TIP.2022.3217695.","journal-title":"IEEE Transactions on Image Processing"},{"key":"1385_CR54","doi-asserted-by":"publisher","first-page":"109555","DOI":"10.1016\/j.patcog.2023.109555","volume":"140","author":"T Zhou","year":"2023","unstructured":"T. Zhou, Y. Zhou, K. L. He, C. Gong, J. Yang, H. Z. Fu, D. G. Shen. Cross-level feature aggregation network for polyp segmentation. Pattern Recognition, vol. 140, pp. 109555, 2023. DOI: https:\/\/doi.org\/10.1016\/j.patcog.2023.109555.","journal-title":"Pattern Recognition"},{"key":"1385_CR55","doi-asserted-by":"publisher","first-page":"106173","DOI":"10.1016\/j.compbiomed.2022.106173","volume":"150","author":"W C Zhang","year":"2022","unstructured":"W. C. Zhang, C. Fu, Y. Zheng, F. Y. Zhang, Y. L. Zhao, C. W. Sham. HSNet: A hybrid semantic network for polyp segmentation. Computers in Biology and Medicine, vol. 150, pp. 106173, 2022. DOI: https:\/\/doi.org\/10.1016\/j.compbiomed.2022.106173.","journal-title":"Computers in Biology and Medicine"},{"key":"1385_CR56","doi-asserted-by":"publisher","unstructured":"X. J. Xiang, Q. Tan, H. Zhou, D. Q. Tang, J. Lai. Multimodal fusion of voice and gesture data for UAV control. Drones, vol. 6, no. 8, Article number 201, 2022. DOI: https:\/\/doi.org\/10.3390\/drones6080201.","DOI":"10.3390\/drones6080201"},{"key":"1385_CR57","doi-asserted-by":"publisher","unstructured":"M. Kaya, H. \u015e. Bilge. Deep metric learning: A survey. Symmetry, vol. 11, no. 9, Article number 1066, 2019. DOI: https:\/\/doi.org\/10.3390\/sym11091066.","DOI":"10.3390\/sym11091066"},{"key":"1385_CR58","unstructured":"W. Ji, J. J. Li, Q. Bi, C. Guo, J. Liu, L. Cheng. Promoting saliency from depth: Deep unsupervised RGB-D saliency detection. In Proceedings of the International Conference on Learning Representations, 2022."},{"key":"1385_CR59","doi-asserted-by":"publisher","unstructured":"Y. R. Piao, W. Ji, J. J. Li, M. Zhang, H. C. Lu. Depth-induced multi-scale recurrent attention network for saliency detection. In Proceedings of the IEEE\/CVF International Conference on Computer Vision, Seoul, Republic of Korea, pp. 7253\u20137262, 2019. DOI: https:\/\/doi.org\/10.1109\/ICCV.2019.00735.","DOI":"10.1109\/ICCV.2019.00735"},{"key":"1385_CR60","doi-asserted-by":"publisher","unstructured":"W. Ji, J. J. Li, C. Bian, Z. C. Zhang, L. Cheng. SemanticRT: A large-scale dataset and method for robust semantic segmentation in multispectral images. In Proceedings of the 31st ACM International Conference on Multimedia, Ottawa, Canada, pp.3307\u20133316, 2023. DOI: https:\/\/doi.org\/10.1145\/3581783.3611738.","DOI":"10.1145\/3581783.3611738"},{"key":"1385_CR61","doi-asserted-by":"publisher","unstructured":"W. Ji, J. J. Li, C. Bian, Z. W. Zhou, J. Y. Zhao, A. Yuille, L. Cheng. Multispectral video semantic segmentation: A benchmark dataset and baseline. In Proceedings of the IEEE\/CVF Conference on Computer Vision and Pattern Recognition, Vancouver, Canada, pp.1094\u20131104, 2023. DOI: https:\/\/doi.org\/10.1109\/CVPR52729.2023.00112.","DOI":"10.1109\/CVPR52729.2023.00112"},{"key":"1385_CR62","doi-asserted-by":"publisher","unstructured":"M. Zhang, J. Liu, Y. F. Wang, Y. R. Piao, S. Y. Yao, W. Ji, J. J. Li, H. C. Lu, Z. X. Luo. Dynamic context-sensitive filtering network for video salient object detection. In Proceedings of the IEEE\/CVF International Conference on Computer Vision, Montreal, Canada, pp. 1533\u20131543, 2021. DOI: https:\/\/doi.org\/10.1109\/ICCV48922.2021.00158.","DOI":"10.1109\/ICCV48922.2021.00158"},{"key":"1385_CR63","doi-asserted-by":"publisher","unstructured":"J. J. Li, T. Y. Yang, W. Ji, J. Wang, L. Cheng. Exploring denoised cross-video contrast for weakly-supervised temporal action localization. In Proceedings of the IEEE\/CVF Conference on Computer Vision and Pattern Recognition, New Orleans, USA, pp. 19882\u201319892, 2022. DOI: https:\/\/doi.org\/10.1109\/CVPR52688.2022.01929.","DOI":"10.1109\/CVPR52688.2022.01929"},{"key":"1385_CR64","unstructured":"J. J. Li, W. Ji, Q. Bi, C. Yan, M. Zhang, Y. R. Piao, H. C. Lu, L. Chen. Joint semantic mining for weakly supervised RGB-D salient object detection. In Proceedings of the 35th Conference on Neural Information Processing Syste, pp. 11945\u201311959, 2021."},{"key":"1385_CR65","doi-asserted-by":"publisher","first-page":"102918","DOI":"10.1016\/j.media.2023.102918","volume":"89","author":"M A Mazurowski","year":"2023","unstructured":"M. A. Mazurowski, H. Y. Dong, H. X. Gu, J. C. Yang, N. Konz, Y. X. Zhang. Segment anything model for medical image analysis: An experimental study. Medical Image Analysis, vol. 89, pp. 102918, 2023. DOI: https:\/\/doi.org\/10.1016\/j.media.2023.102918.","journal-title":"Medical Image Analysis"},{"key":"1385_CR66","doi-asserted-by":"crossref","unstructured":"Y. C. Zhang, R. S. Jiao. How segment anything model (SAM) boost medical image segmentation? [Online], Available: https:\/\/arxiv.org\/abs\/2305.03678, 2023.","DOI":"10.2139\/ssrn.4495221"},{"key":"1385_CR67","unstructured":"J. Ma, B. Wang. Segment anything in medical images, [Online], Available: https:\/\/arxiv.org\/abs\/2304.12306, 2023."},{"key":"1385_CR68","unstructured":"J. D. Wu, R. Fu, H. H. Fang, Y. P. Liu, Z. W. Wang, Y. W. Xu, Y. M. Jin, T. Arbel. Medical SAM adapter: Adapting segment anything model for medical image segmentation, [Online], Available: https:\/\/arxiv.org\/abs\/2304.12620, 2023."},{"key":"1385_CR69","doi-asserted-by":"publisher","unstructured":"L. P. Osco, Q. S. Wu, E. L. De Lemos, W. N. Gon\u00e7alves, A. P. M. Ramos, J. Li, J. M. Junior. The segment anything model (SAM) for remote sensing applications: From zero to one shot. International Journal of Applied Earth Observation and Geoinformation, vol. 124, Article number 103540, 2023. DOI: https:\/\/doi.org\/10.1016\/j.jag.2023.103540.","DOI":"10.1016\/j.jag.2023.103540"},{"key":"1385_CR70","doi-asserted-by":"crossref","unstructured":"F. Chen, M. V. Giuffrida, S. A. Tsaftaris. Adapting vision foundation models for plant phenotyping. In Proceedings of the IEEE\/CVF International Conference on Computer Vision Workshop, pp. 604\u2013613, 2023.","DOI":"10.1109\/ICCVW60793.2023.00067"},{"key":"1385_CR71","doi-asserted-by":"publisher","unstructured":"Y. Zhao, K. C. Song, W. Q. Cui, H. Ren, Y. H. Yan. MFS enhanced SAM: Achieving superior performance in bimodal few-shot segmentation. Journal of Visual Communication and Image Representation, vo1. 97, Article number 103946, 2023. DOI: https:\/\/doi.org\/10.1016\/j.jvcir.2023.103946.","DOI":"10.1016\/j.jvcir.2023.103946"},{"key":"1385_CR72","unstructured":"Y. M. Cheng, L. L. Li, Y. Y. Xu, X. D. Li, Z. X. Yang, W. G. Wang, Y. Yang. Segment and track anything, [Online], Available: https:\/\/arxiv.org\/abs\/2305.06558, 2023."},{"key":"1385_CR73","unstructured":"Z. H. Lu, Z. Y. Xiao, J. W. Bai, Z. W. Xiong, X. C. Wang. Can sam boost video super-resolutionn [Online], Available: https:\/\/arxiv.org\/abs\/2305.06524, 2023."},{"key":"1385_CR74","doi-asserted-by":"crossref","unstructured":"T. R. Chen, L. Y. Zhu, C. T. Deng, R. L. Cao, Y. Wang, S. Z. Zhang, Z. J. Li, L. Y. Sun, Y. Zang, P. P. Mao. SAM-adapter: Adapting segment anything in underperformed scenes. In Proceedings of the IEEE\/CVF International Conference on Computer Vision Workshop, pp. 3367\u20133375, 2023.","DOI":"10.1109\/ICCVW60793.2023.00361"},{"key":"1385_CR75","unstructured":"H. X. Dai, C. Ma, Z. L. Liu, Y. W. Li, P. Shu, X. Z. Wei, L. Zhao, Z. H. Wu, D. J. Zhu, W. Liu, Q. Z. Li, T. M. Liu, X. Li. SAMAug: Point prompt augmentation for segment anything model, [Online], Available: https:\/\/arxiv.org\/abs\/2307.01187, 2023."}],"updated-by":[{"DOI":"10.1007\/s11633-024-1526-0","type":"correction","label":"Correction","source":"publisher","updated":{"date-parts":[[2024,9,11]],"date-time":"2024-09-11T00:00:00Z","timestamp":1726012800000}}],"container-title":["Machine Intelligence Research"],"original-title":[],"language":"en","link":[{"URL":"https:\/\/link.springer.com\/content\/pdf\/10.1007\/s11633-023-1385-0.pdf","content-type":"application\/pdf","content-version":"vor","intended-application":"text-mining"},{"URL":"https:\/\/link.springer.com\/article\/10.1007\/s11633-023-1385-0","content-type":"text\/html","content-version":"vor","intended-application":"text-mining"},{"URL":"https:\/\/link.springer.com\/content\/pdf\/10.1007\/s11633-023-1385-0.pdf","content-type":"application\/pdf","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2026,6,4]],"date-time":"2026-06-04T08:31:34Z","timestamp":1780561894000},"score":1,"resource":{"primary":{"URL":"https:\/\/link.springer.com\/10.1007\/s11633-023-1385-0"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2024,4,12]]},"references-count":75,"journal-issue":{"issue":"4","published-print":{"date-parts":[[2024,8]]}},"alternative-id":["1385"],"URL":"https:\/\/doi.org\/10.1007\/s11633-023-1385-0","relation":{},"ISSN":["2731-538X","2731-5398"],"issn-type":[{"value":"2731-538X","type":"print"},{"value":"2731-5398","type":"electronic"}],"subject":[],"published":{"date-parts":[[2024,4,12]]}}}