{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2026,5,15]],"date-time":"2026-05-15T13:40:54Z","timestamp":1778852454817,"version":"3.51.4"},"reference-count":74,"publisher":"Association for Computing Machinery (ACM)","issue":"2","license":[{"start":{"date-parts":[[2025,5,26]],"date-time":"2025-05-26T00:00:00Z","timestamp":1748217600000},"content-version":"vor","delay-in-days":0,"URL":"https:\/\/www.acm.org\/publications\/policies\/copyright_policy#Background"}],"content-domain":{"domain":["dl.acm.org"],"crossmark-restriction":true},"short-container-title":["Proc. ACM Comput. Graph. Interact. Tech."],"published-print":{"date-parts":[[2025,6]]},"abstract":"<jats:p>We explore the transformative potential of SAM 2, a vision foundation model, in advancing gaze estimation. SAM 2 addresses key challenges in gaze estimation by significantly reducing annotation time, simplifying deployment, and enhancing segmentation accuracy. Utilizing its zero-shot capabilities with minimal user input\u2014a single click per video\u2014we tested SAM 2 on over 14 million eye images from a diverse range of datasets, including the EDS challenge datasets and Labelled Pupils in the Wild. This is the first application of SAM 2 to the gaze estimation domain. Remarkably, SAM 2 matches the performance of domain-specific models in pupil segmentation, achieving competitive mIOU scores of up to 93% without fine-tuning. We argue that SAM 2 achieves the sought-after standard of domain generalization, with consistent mIOU scores (89.71%-93.74%) across diverse datasets, from virtual reality to \"gaze-in-the-wild\" scenarios. We provide our code and segmentation masks for these datasets to promote further research.<\/jats:p>","DOI":"10.1145\/3729409","type":"journal-article","created":{"date-parts":[[2025,5,27]],"date-time":"2025-05-27T05:51:27Z","timestamp":1748325087000},"page":"1-16","update-policy":"https:\/\/doi.org\/10.1145\/crossmark-policy","source":"Crossref","is-referenced-by-count":3,"title":["Zero-Shot Pupil Segmentation with SAM 2: A Case Study of Over 14 Million Images"],"prefix":"10.1145","volume":"8","author":[{"ORCID":"https:\/\/orcid.org\/0009-0005-5580-2304","authenticated-orcid":false,"given":"Virmarie","family":"Maquiling","sequence":"first","affiliation":[{"name":"Human-Centered Technologies for Learning, Technical University of Munich, Munich, Bavaria, Germany"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"ORCID":"https:\/\/orcid.org\/0009-0004-5685-7318","authenticated-orcid":false,"given":"Sean Anthony","family":"Byrne","sequence":"additional","affiliation":[{"name":"Dipartimento di Elettronica, Informazione e Bioingegneria, Politecnico di Milano, Milan, Italy"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"ORCID":"https:\/\/orcid.org\/0000-0002-4672-8756","authenticated-orcid":false,"given":"Diederick C.","family":"Niehorster","sequence":"additional","affiliation":[{"name":"Lund University Humanities Lab, Lund University, Lund, Sweden, and Department of Psychology, Lund University, Lund, Sweden"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"ORCID":"https:\/\/orcid.org\/0000-0002-3485-4317","authenticated-orcid":false,"given":"Marco","family":"Carminati","sequence":"additional","affiliation":[{"name":"DEIB, Politecnico di Milano, Milano, MI, Italy"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"ORCID":"https:\/\/orcid.org\/0000-0003-3146-4484","authenticated-orcid":false,"given":"Enkelejda","family":"Kasneci","sequence":"additional","affiliation":[{"name":"Human-Centered Technologies for Learning, Technical University of Munich, Munich, Germany"}],"role":[{"role":"author","vocabulary":"crossref"}]}],"member":"320","published-online":{"date-parts":[[2025,5,26]]},"reference":[{"key":"e_1_3_1_2_1","article-title":"Gpt-4 technical report","author":"Achiam Josh","year":"2023","unstructured":"Josh Achiam, Steven Adler, Sandhini Agarwal, Lama Ahmad, Ilge Akkaya, Florencia\u00a0Leoni Aleman, Diogo Almeida, Janko Altenschmidt, Sam Altman, Shyamal Anadkat, et\u00a0al. 2023. Gpt-4 technical report. arXiv preprint arXiv:https:\/\/arXiv.org\/abs\/2303.08774 (2023).","journal-title":"arXiv preprint arXiv:https:\/\/arXiv.org\/abs\/2303.08774"},{"key":"e_1_3_1_3_1","doi-asserted-by":"publisher","DOI":"10.1109\/ICIP.2006.312773"},{"key":"e_1_3_1_4_1","article-title":"On the opportunities and risks of foundation models","author":"Bommasani Rishi","year":"2021","unstructured":"Rishi Bommasani, Drew\u00a0A Hudson, Ehsan Adeli, Russ Altman, Simran Arora, Sydney von Arx, Michael\u00a0S Bernstein, Jeannette Bohg, Antoine Bosselut, Emma Brunskill, et\u00a0al. 2021. On the opportunities and risks of foundation models. arXiv preprint arXiv:https:\/\/arXiv.org\/abs\/2108.07258 (2021).","journal-title":"arXiv preprint arXiv:https:\/\/arXiv.org\/abs\/2108.07258"},{"key":"e_1_3_1_5_1","doi-asserted-by":"publisher","DOI":"10.1145\/3379156.3391364"},{"key":"e_1_3_1_6_1","unstructured":"Tom\u00a0B Brown. 2020. Language models are few-shot learners. arXiv preprint arXiv:https:\/\/arXiv.org\/abs\/2005.14165 (2020)."},{"key":"e_1_3_1_7_1","doi-asserted-by":"publisher","DOI":"10.1109\/AIxVR59861.2024.00020"},{"key":"e_1_3_1_8_1","unstructured":"Sean\u00a0Anthony Byrne Virmarie Maquiling Marcus Nystr\u00f6m Enkelejda Kasneci and Diederick\u00a0C. Niehorster. 2023. LEyes: A Lightweight Framework for Deep Learning-Based Eye Tracking using Synthetic Eye Images. arxiv:https:\/\/arXiv.org\/abs\/2309.06129\u00a0[cs.CV] https:\/\/arxiv.org\/abs\/2309.06129"},{"key":"e_1_3_1_9_1","doi-asserted-by":"publisher","DOI":"10.3758\/s13428-023-02297-w"},{"key":"e_1_3_1_10_1","first-page":"25971","article-title":"Segment anything in 3d with nerfs","volume":"36","author":"Cen Jiazhong","year":"2023","unstructured":"Jiazhong Cen, Zanwei Zhou, Jiemin Fang, Wei Shen, Lingxi Xie, Dongsheng Jiang, Xiaopeng Zhang, Qi Tian, et\u00a0al. 2023. Segment anything in 3d with nerfs. Advances in Neural Information Processing Systems 36 (2023), 25971\u201325990.","journal-title":"Advances in Neural Information Processing Systems"},{"key":"e_1_3_1_11_1","doi-asserted-by":"publisher","DOI":"10.1109\/ICCVW.2019.00568"},{"key":"e_1_3_1_12_1","article-title":"Segment and track anything","author":"Cheng Yangming","year":"2023","unstructured":"Yangming Cheng, Liulei Li, Yuanyou Xu, Xiaodi Li, Zongxin Yang, Wenguan Wang, and Yi Yang. 2023. Segment and track anything. arXiv preprint arXiv:https:\/\/arXiv.org\/abs\/2305.06558 (2023).","journal-title":"arXiv preprint arXiv:https:\/\/arXiv.org\/abs\/2305.06558"},{"key":"e_1_3_1_13_1","unstructured":"Jiangfan Deng Zhuang Jia Zhaoxue Wang Xiang Long and Daniel\u00a0K Du. 2024. Towards Unsupervised Eye-Region Segmentation for Eye Tracking. arXiv preprint arXiv:https:\/\/arXiv.org\/abs\/2410.06131 (2024)."},{"key":"e_1_3_1_14_1","article-title":"Adapting segment anything model for change detection in VHR remote sensing images","author":"Ding Lei","year":"2024","unstructured":"Lei Ding, Kun Zhu, Daifeng Peng, Hao Tang, Kuiwu Yang, and Lorenzo Bruzzone. 2024. Adapting segment anything model for change detection in VHR remote sensing images. IEEE Transactions on Geoscience and Remote Sensing (2024).","journal-title":"IEEE Transactions on Geoscience and Remote Sensing"},{"key":"e_1_3_1_15_1","doi-asserted-by":"publisher","DOI":"10.1145\/3204493.3204559"},{"key":"e_1_3_1_16_1","doi-asserted-by":"publisher","DOI":"10.1109\/ISMAR52148.2021.00053"},{"key":"e_1_3_1_17_1","doi-asserted-by":"publisher","DOI":"10.1007\/978-3-319-23192-1_4"},{"key":"e_1_3_1_18_1","doi-asserted-by":"publisher","DOI":"10.1109\/WACV.2017.126"},{"key":"e_1_3_1_19_1","article-title":"Pupilnet: Convolutional neural networks for robust pupil detection","author":"Fuhl Wolfgang","year":"2016","unstructured":"Wolfgang Fuhl, Thiago Santini, Gjergji Kasneci, and Enkelejda Kasneci. 2016a. Pupilnet: Convolutional neural networks for robust pupil detection. arXiv preprint arXiv:https:\/\/arXiv.org\/abs\/1601.04902 (2016).","journal-title":"arXiv preprint arXiv:https:\/\/arXiv.org\/abs\/1601.04902"},{"key":"e_1_3_1_20_1","article-title":"Pupilnet v2. 0: Convolutional neural networks for cpu based real time robust pupil detection","author":"Fuhl Wolfgang","year":"2017","unstructured":"Wolfgang Fuhl, Thiago Santini, Gjergji Kasneci, Wolfgang Rosenstiel, and Enkelejda Kasneci. 2017b. Pupilnet v2. 0: Convolutional neural networks for cpu based real time robust pupil detection. arXiv preprint arXiv:https:\/\/arXiv.org\/abs\/1711.00112 (2017).","journal-title":"arXiv preprint arXiv:https:\/\/arXiv.org\/abs\/1711.00112"},{"key":"e_1_3_1_21_1","doi-asserted-by":"publisher","DOI":"10.1145\/2857491.2857505"},{"key":"e_1_3_1_22_1","article-title":"Openeds: Open eye dataset","author":"Garbin Stephan\u00a0J","year":"2019","unstructured":"Stephan\u00a0J Garbin, Yiru Shen, Immo Schuetz, Robert Cavin, Gregory Hughes, and Sachin\u00a0S Talathi. 2019. Openeds: Open eye dataset. arXiv preprint arXiv:https:\/\/arXiv.org\/abs\/1905.03702 (2019).","journal-title":"arXiv preprint arXiv:https:\/\/arXiv.org\/abs\/1905.03702"},{"key":"e_1_3_1_23_1","article-title":"TimeGPT-1","author":"Garza Azul","year":"2023","unstructured":"Azul Garza, Cristian Challu, and Max Mergenthaler-Canseco. 2023. TimeGPT-1. arXiv preprint arXiv:https:\/\/arXiv.org\/abs\/2310.03589 (2023).","journal-title":"arXiv preprint arXiv:https:\/\/arXiv.org\/abs\/2310.03589"},{"key":"e_1_3_1_24_1","article-title":"Sam2point: Segment any 3d as videos in zero-shot and promptable manners","author":"Guo Ziyu","year":"2024","unstructured":"Ziyu Guo, Renrui Zhang, Xiangyang Zhu, Chengzhuo Tong, Peng Gao, Chunyuan Li, and Pheng-Ann Heng. 2024. Sam2point: Segment any 3d as videos in zero-shot and promptable manners. arXiv preprint arXiv:https:\/\/arXiv.org\/abs\/2408.16768 (2024).","journal-title":"arXiv preprint arXiv:https:\/\/arXiv.org\/abs\/2408.16768"},{"key":"e_1_3_1_25_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR.2016.90"},{"key":"e_1_3_1_26_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR.2017.243"},{"key":"e_1_3_1_27_1","doi-asserted-by":"publisher","DOI":"10.1016\/j.media.2023.103061"},{"key":"e_1_3_1_28_1","article-title":"CondSeg: Ellipse Estimation of Pupil and Iris via Conditioned Segmentation","author":"Jia Zhuang","year":"2024","unstructured":"Zhuang Jia, Jiangfan Deng, Liying Chi, Xiang Long, and Daniel\u00a0K Du. 2024. CondSeg: Ellipse Estimation of Pupil and Iris via Conditioned Segmentation. arXiv preprint arXiv:https:\/\/arXiv.org\/abs\/2408.17231 (2024).","journal-title":"arXiv preprint arXiv:https:\/\/arXiv.org\/abs\/2408.17231"},{"key":"e_1_3_1_29_1","first-page":"23082","article-title":"Position: On the societal impact of open foundation models","volume":"235","author":"Kapoor Sayash","year":"2024","unstructured":"Sayash Kapoor, Rishi Bommasani, Kevin Klyman, Shayne Longpre, Ashwin Ramaswami, Peter Cihon, Aspen Hopkins, Kevin Bankston, Stella Biderman, Miranda Bogen, et\u00a0al. 2024. Position: On the societal impact of open foundation models. Proc. Machine Learning Res 235 (2024), 23082\u201323104.","journal-title":"Proc. Machine Learning Res"},{"key":"e_1_3_1_30_1","doi-asserted-by":"publisher","DOI":"10.1371\/journal.pone.0087470"},{"key":"e_1_3_1_31_1","first-page":"2","volume-title":"Proceedings of naacL-HLT","author":"Kenton Jacob Devlin Ming-Wei\u00a0Chang","year":"2019","unstructured":"Jacob Devlin Ming-Wei\u00a0Chang Kenton and Lee\u00a0Kristina Toutanova. 2019. Bert: Pre-training of deep bidirectional transformers for language understanding. In Proceedings of naacL-HLT, Vol.\u00a01. Minneapolis, Minnesota, 2."},{"key":"e_1_3_1_32_1","doi-asserted-by":"publisher","DOI":"10.1145\/3290605.3300780"},{"key":"e_1_3_1_33_1","doi-asserted-by":"publisher","DOI":"10.1109\/ICCV51070.2023.00371"},{"key":"e_1_3_1_34_1","doi-asserted-by":"publisher","DOI":"10.1007\/978-3-031-72390-2_49"},{"key":"e_1_3_1_35_1","doi-asserted-by":"publisher","DOI":"10.1038\/s41598-020-59251-5"},{"key":"e_1_3_1_36_1","doi-asserted-by":"publisher","DOI":"10.1145\/3530880"},{"key":"e_1_3_1_37_1","doi-asserted-by":"publisher","DOI":"10.1109\/TVCG.2021.3067765"},{"key":"e_1_3_1_38_1","doi-asserted-by":"publisher","DOI":"10.1038\/s41467-024-44824-z"},{"key":"e_1_3_1_39_1","doi-asserted-by":"publisher","DOI":"10.1145\/3654704"},{"key":"e_1_3_1_40_1","doi-asserted-by":"publisher","DOI":"10.1016\/j.media.2023.102918"},{"key":"e_1_3_1_41_1","doi-asserted-by":"publisher","DOI":"10.1109\/ICIP.2009.5414530"},{"key":"e_1_3_1_42_1","doi-asserted-by":"publisher","DOI":"10.1145\/3385955.3407935"},{"key":"e_1_3_1_43_1","doi-asserted-by":"publisher","DOI":"10.1145\/3654703"},{"key":"e_1_3_1_44_1","doi-asserted-by":"publisher","DOI":"10.3758\/s13428-021-01780-6"},{"key":"e_1_3_1_45_1","article-title":"Openeds2020: Open eyes dataset","author":"Palmero Cristina","year":"2020","unstructured":"Cristina Palmero, Abhishek Sharma, Karsten Behrendt, Kapil Krishnakumar, Oleg\u00a0V Komogortsev, and Sachin\u00a0S Talathi. 2020. Openeds2020: Open eyes dataset. arXiv preprint arXiv:https:\/\/arXiv.org\/abs\/2005.03876 (2020).","journal-title":"arXiv preprint arXiv:https:\/\/arXiv.org\/abs\/2005.03876"},{"key":"e_1_3_1_46_1","unstructured":"Antonio Per\u00e9z M\u00a0Luisa C\u00f3rdoba A Garcia Rafael M\u00e9ndez ML Munoz Jos\u00e9\u00a0Luis Pedraza and F Sanchez. 2003. A precise eye-gaze detection and tracking system. (2003)."},{"key":"e_1_3_1_47_1","doi-asserted-by":"publisher","DOI":"10.1109\/ACCESS.2023.3315235"},{"key":"e_1_3_1_48_1","unstructured":"Alec Radford. 2018. Improving language understanding by generative pre-training. (2018)."},{"issue":"8","key":"e_1_3_1_49_1","first-page":"9","article-title":"Language models are unsupervised multitask learners","volume":"1","author":"Radford Alec","year":"2019","unstructured":"Alec Radford, Jeffrey Wu, Rewon Child, David Luan, Dario Amodei, Ilya Sutskever, et\u00a0al. 2019. Language models are unsupervised multitask learners. OpenAI blog 1, 8 (2019), 9.","journal-title":"OpenAI blog"},{"key":"e_1_3_1_50_1","article-title":"Sam 2: Segment anything in images and videos","author":"Ravi Nikhila","year":"2024","unstructured":"Nikhila Ravi, Valentin Gabeur, Yuan-Ting Hu, Ronghang Hu, Chaitanya Ryali, Tengyu Ma, Haitham Khedr, Roman R\u00e4dle, Chloe Rolland, Laura Gustafson, et\u00a0al. 2024. Sam 2: Segment anything in images and videos. arXiv preprint arXiv:https:\/\/arXiv.org\/abs\/2408.00714 (2024).","journal-title":"arXiv preprint arXiv:https:\/\/arXiv.org\/abs\/2408.00714"},{"key":"e_1_3_1_51_1","doi-asserted-by":"publisher","DOI":"10.1007\/978-3-319-24574-4_28"},{"key":"e_1_3_1_52_1","doi-asserted-by":"publisher","DOI":"10.1145\/3411764.3445518"},{"key":"e_1_3_1_53_1","doi-asserted-by":"publisher","DOI":"10.1016\/j.cviu.2018.02.002"},{"key":"e_1_3_1_54_1","doi-asserted-by":"publisher","DOI":"10.1145\/3204493.3204578"},{"key":"e_1_3_1_55_1","first-page":"2024","article-title":"Nicheformer: a foundation model for single-cell and spatial omics","author":"Schaar Anna\u00a0Christina","year":"2024","unstructured":"Anna\u00a0Christina Schaar, Alejandro Tejada-Lapuerta, Giovanni Palla, Robert Gutgesell, Lennard Halle, Mariia Minaeva, Larsen Vornholz, Leander Dony, Francesca Drummer, Mojtaba Bahrami, et\u00a0al. 2024. Nicheformer: a foundation model for single-cell and spatial omics. bioRxiv (2024), 2024\u201304.","journal-title":"bioRxiv"},{"key":"e_1_3_1_56_1","first-page":"1","article-title":"Semantic segmentation of glaciological features across multiple remote sensing platforms with the Segment Anything Model (SAM)","author":"Shankar Siddharth","year":"2023","unstructured":"Siddharth Shankar, Leigh\u00a0A Stearns, and CJ van\u00a0der Veen. 2023. Semantic segmentation of glaciological features across multiple remote sensing platforms with the Segment Anything Model (SAM). Journal of Glaciology (2023), 1\u201310.","journal-title":"Journal of Glaciology"},{"key":"e_1_3_1_57_1","article-title":"Anything-3d: Towards single-view anything reconstruction in the wild","author":"Shen Qiuhong","year":"2023","unstructured":"Qiuhong Shen, Xingyi Yang, and Xinchao Wang. 2023. Anything-3d: Towards single-view anything reconstruction in the wild. arXiv preprint arXiv:https:\/\/arXiv.org\/abs\/2304.10261 (2023).","journal-title":"arXiv preprint arXiv:https:\/\/arXiv.org\/abs\/2304.10261"},{"key":"e_1_3_1_58_1","doi-asserted-by":"publisher","DOI":"10.1117\/12.189136"},{"key":"e_1_3_1_59_1","doi-asserted-by":"publisher","DOI":"10.1109\/ICCV.2017.97"},{"key":"e_1_3_1_60_1","doi-asserted-by":"publisher","DOI":"10.1145\/2168556.2168585"},{"key":"e_1_3_1_61_1","doi-asserted-by":"publisher","DOI":"10.1145\/2857491.2857520"},{"key":"e_1_3_1_62_1","article-title":"Samrs: Scaling-up remote sensing segmentation dataset with segment anything model","volume":"36","author":"Wang Di","year":"2024","unstructured":"Di Wang, Jing Zhang, Bo Du, Minqiang Xu, Lin Liu, Dacheng Tao, and Liangpei Zhang. 2024. Samrs: Scaling-up remote sensing segmentation dataset with segment anything model. Advances in Neural Information Processing Systems 36 (2024).","journal-title":"Advances in Neural Information Processing Systems"},{"key":"e_1_3_1_63_1","doi-asserted-by":"publisher","DOI":"10.1109\/IAVVC63304.2024.10786458"},{"key":"e_1_3_1_64_1","article-title":"Track anything: Segment anything meets videos","author":"Yang Jinyu","year":"2023","unstructured":"Jinyu Yang, Mingqi Gao, Zhe Li, Shang Gao, Fangjing Wang, and Feng Zheng. 2023a. Track anything: Segment anything meets videos. arXiv preprint arXiv:https:\/\/arXiv.org\/abs\/2304.11968 (2023).","journal-title":"arXiv preprint arXiv:https:\/\/arXiv.org\/abs\/2304.11968"},{"key":"e_1_3_1_65_1","article-title":"Sam3d: Segment anything in 3d scenes","author":"Yang Yunhan","year":"2023","unstructured":"Yunhan Yang, Xiaoyang Wu, Tong He, Hengshuang Zhao, and Xihui Liu. 2023b. Sam3d: Segment anything in 3d scenes. arXiv preprint arXiv:https:\/\/arXiv.org\/abs\/2306.03908 (2023).","journal-title":"arXiv preprint arXiv:https:\/\/arXiv.org\/abs\/2306.03908"},{"key":"e_1_3_1_66_1","doi-asserted-by":"publisher","DOI":"10.1016\/j.jneumeth.2019.05.016"},{"key":"e_1_3_1_67_1","article-title":"Inpaint anything: Segment anything meets image inpainting","author":"Yu Tao","year":"2023","unstructured":"Tao Yu, Runseng Feng, Ruoyu Feng, Jinming Liu, Xin Jin, Wenjun Zeng, and Zhibo Chen. 2023. Inpaint anything: Segment anything meets image inpainting. arXiv preprint arXiv:https:\/\/arXiv.org\/abs\/2304.06790 (2023).","journal-title":"arXiv preprint arXiv:https:\/\/arXiv.org\/abs\/2304.06790"},{"key":"e_1_3_1_68_1","article-title":"Faster segment anything: Towards lightweight sam for mobile applications","author":"Zhang Chaoning","year":"2023","unstructured":"Chaoning Zhang, Dongshen Han, Yu Qiao, Jung\u00a0Uk Kim, Sung-Ho Bae, Seungkyu Lee, and Choong\u00a0Seon Hong. 2023a. Faster segment anything: Towards lightweight sam for mobile applications. arXiv preprint arXiv:https:\/\/arXiv.org\/abs\/2306.14289 (2023).","journal-title":"arXiv preprint arXiv:https:\/\/arXiv.org\/abs\/2306.14289"},{"key":"e_1_3_1_69_1","article-title":"Mobilesamv2: Faster segment anything to everything","author":"Zhang Chaoning","year":"2023","unstructured":"Chaoning Zhang, Dongshen Han, Sheng Zheng, Jinwoo Choi, Tae-Ho Kim, and Choong\u00a0Seon Hong. 2023b. Mobilesamv2: Faster segment anything to everything. arXiv preprint arXiv:https:\/\/arXiv.org\/abs\/2312.09579 (2023).","journal-title":"arXiv preprint arXiv:https:\/\/arXiv.org\/abs\/2312.09579"},{"key":"e_1_3_1_70_1","doi-asserted-by":"publisher","DOI":"10.1016\/j.media.2023.102996"},{"key":"e_1_3_1_71_1","doi-asserted-by":"publisher","DOI":"10.1016\/j.compbiomed.2024.108238"},{"key":"e_1_3_1_72_1","article-title":"Fast segment anything","author":"Zhao Xu","year":"2023","unstructured":"Xu Zhao, Wenchao Ding, Yongqi An, Yinglong Du, Tao Yu, Min Li, Ming Tang, and Jinqiao Wang. 2023. Fast segment anything. arXiv preprint arXiv:https:\/\/arXiv.org\/abs\/2306.12156 (2023).","journal-title":"arXiv preprint arXiv:https:\/\/arXiv.org\/abs\/2306.12156"},{"key":"e_1_3_1_73_1","doi-asserted-by":"publisher","DOI":"10.1109\/ICDSCA59871.2023.10393594"},{"key":"e_1_3_1_74_1","doi-asserted-by":"publisher","DOI":"10.1038\/s41586-023-06555-x"},{"key":"e_1_3_1_75_1","article-title":"Medical sam 2: Segment medical images as video via segment anything model 2","author":"Zhu Jiayuan","year":"2024","unstructured":"Jiayuan Zhu, Yunli Qi, and Junde Wu. 2024. Medical sam 2: Segment medical images as video via segment anything model 2. arXiv preprint arXiv:https:\/\/arXiv.org\/abs\/2408.00874 (2024).","journal-title":"arXiv preprint arXiv:https:\/\/arXiv.org\/abs\/2408.00874"}],"container-title":["Proceedings of the ACM on Computer Graphics and Interactive Techniques"],"original-title":[],"language":"en","link":[{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/3729409","content-type":"unspecified","content-version":"vor","intended-application":"text-mining"}],"deposited":{"date-parts":[[2025,6,19]],"date-time":"2025-06-19T01:56:57Z","timestamp":1750298217000},"score":1,"resource":{"primary":{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/3729409"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2025,5,26]]},"references-count":74,"journal-issue":{"issue":"2","published-print":{"date-parts":[[2025,6]]}},"alternative-id":["10.1145\/3729409"],"URL":"https:\/\/doi.org\/10.1145\/3729409","relation":{},"ISSN":["2577-6193"],"issn-type":[{"value":"2577-6193","type":"electronic"}],"subject":[],"published":{"date-parts":[[2025,5,26]]},"assertion":[{"value":"2025-05-26","order":3,"name":"published","label":"Published","group":{"name":"publication_history","label":"Publication History"}}]}}