{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2025,6,18]],"date-time":"2025-06-18T04:17:09Z","timestamp":1750220229620,"version":"3.41.0"},"publisher-location":"New York, NY, USA","reference-count":35,"publisher":"ACM","license":[{"start":{"date-parts":[[2022,6,27]],"date-time":"2022-06-27T00:00:00Z","timestamp":1656288000000},"content-version":"vor","delay-in-days":0,"URL":"https:\/\/www.acm.org\/publications\/policies\/copyright_policy#Background"}],"funder":[{"name":"National Natural Science Foundation of China","award":["61972333"],"award-info":[{"award-number":["61972333"]}]},{"name":"Research Foundation of Education Department of Hunan Province of China","award":["21B0172"],"award-info":[{"award-number":["21B0172"]}]},{"DOI":"10.13039\/501100001809","name":"National Natural Science Foundation of China","doi-asserted-by":"publisher","award":["61802328"],"award-info":[{"award-number":["61802328"]}],"id":[{"id":"10.13039\/501100001809","id-type":"DOI","asserted-by":"publisher"}]}],"content-domain":{"domain":["dl.acm.org"],"crossmark-restriction":true},"short-container-title":[],"published-print":{"date-parts":[[2022,6,27]]},"DOI":"10.1145\/3512527.3531420","type":"proceedings-article","created":{"date-parts":[[2022,6,23]],"date-time":"2022-06-23T22:23:32Z","timestamp":1656023012000},"page":"654-660","update-policy":"https:\/\/doi.org\/10.1145\/crossmark-policy","source":"Crossref","is-referenced-by-count":1,"title":["I2-Net: Intra- and Inter-scale Collaborative Learning Network for Abdominal Multi-organ Segmentation"],"prefix":"10.1145","author":[{"given":"Chao","family":"Suo","sequence":"first","affiliation":[{"name":"Xiangtan University, Xiangtan, China"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Xuanya","family":"Li","sequence":"additional","affiliation":[{"name":"Baidu Inc., Beijing, China"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Donghui","family":"Tan","sequence":"additional","affiliation":[{"name":"Affiliated Hospital of Xiangnan University, Chenzhou, China"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Yuan","family":"Zhang","sequence":"additional","affiliation":[{"name":"Xiangtan University, Xiangtan, China"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Xieping","family":"Gao","sequence":"additional","affiliation":[{"name":"Hunan Normal University, Changsha, China"}],"role":[{"role":"author","vocabulary":"crossref"}]}],"member":"320","published-online":{"date-parts":[[2022,6,27]]},"reference":[{"key":"e_1_3_2_2_1_1","volume-title":"Swin-Unet: Unet-like Pure Transformer for Medical Image Segmentation. arXiv preprint arXiv:2105.05537","author":"Cao Hu","year":"2021","unstructured":"Hu Cao , Yueyue Wang , Joy Chen , Dongsheng Jiang , Xiaopeng Zhang , Qi Tian , and Manning Wang . 2021. Swin-Unet: Unet-like Pure Transformer for Medical Image Segmentation. arXiv preprint arXiv:2105.05537 ( 2021 ). Hu Cao, Yueyue Wang, Joy Chen, Dongsheng Jiang, Xiaopeng Zhang, Qi Tian, and Manning Wang. 2021. Swin-Unet: Unet-like Pure Transformer for Medical Image Segmentation. arXiv preprint arXiv:2105.05537 (2021)."},{"key":"e_1_3_2_2_2_1","volume-title":"Transclaw u-net: Claw u-net with transformers for medical image segmentation. arXiv preprint arXiv:2107.05188","author":"Chang Yao","year":"2021","unstructured":"Yao Chang , Hu Menghan , Zhai Guangtao , and Zhang Xiao-Ping . 2021. Transclaw u-net: Claw u-net with transformers for medical image segmentation. arXiv preprint arXiv:2107.05188 ( 2021 ). Yao Chang, Hu Menghan, Zhai Guangtao, and Zhang Xiao-Ping. 2021. Transclaw u-net: Claw u-net with transformers for medical image segmentation. arXiv preprint arXiv:2107.05188 (2021)."},{"key":"e_1_3_2_2_3_1","volume-title":"Transunet: Transformers make strong encoders for medical image segmentation. arXiv preprint arXiv:2102.04306","author":"Chen Jieneng","year":"2021","unstructured":"Jieneng Chen , Yongyi Lu , Qihang Yu , Xiangde Luo , Ehsan Adeli , Yan Wang , Le Lu , Alan L Yuille , and Yuyin Zhou . 2021 . Transunet: Transformers make strong encoders for medical image segmentation. arXiv preprint arXiv:2102.04306 (2021). Jieneng Chen, Yongyi Lu, Qihang Yu, Xiangde Luo, Ehsan Adeli, Yan Wang, Le Lu, Alan L Yuille, and Yuyin Zhou. 2021. Transunet: Transformers make strong encoders for medical image segmentation. arXiv preprint arXiv:2102.04306 (2021)."},{"key":"e_1_3_2_2_4_1","volume-title":"Convit: Improving vision transformers with soft convolutional inductive biases. arXiv preprint arXiv:2103.10697","author":"Ascoli St\u00e9phane","year":"2021","unstructured":"St\u00e9phane d' Ascoli , Hugo Touvron , Matthew Leavitt , Ari Morcos , Giulio Biroli , and Levent Sagun . 2021 . Convit: Improving vision transformers with soft convolutional inductive biases. arXiv preprint arXiv:2103.10697 (2021). St\u00e9phane d'Ascoli, Hugo Touvron, Matthew Leavitt, Ari Morcos, Giulio Biroli, and Levent Sagun. 2021. Convit: Improving vision transformers with soft convolutional inductive biases. arXiv preprint arXiv:2103.10697 (2021)."},{"key":"e_1_3_2_2_5_1","volume-title":"An Image is Worth 16x16 Words: Transformers for Image Recognition at Scale. ICLR","author":"Dosovitskiy Alexey","year":"2021","unstructured":"Alexey Dosovitskiy , Lucas Beyer , Alexander Kolesnikov , Dirk Weissenborn , Xiaohua Zhai , Thomas Unterthiner , Mostafa Dehghani , Matthias Minderer , Georg Heigold , Sylvain Gelly , Jakob Uszkoreit , and Neil Houlsby . 2021. An Image is Worth 16x16 Words: Transformers for Image Recognition at Scale. ICLR ( 2021 ). Alexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn, Xiaohua Zhai, Thomas Unterthiner, Mostafa Dehghani, Matthias Minderer, Georg Heigold, Sylvain Gelly, Jakob Uszkoreit, and Neil Houlsby. 2021. An Image is Worth 16x16 Words: Transformers for Image Recognition at Scale. ICLR (2021)."},{"doi-asserted-by":"publisher","key":"e_1_3_2_2_6_1","DOI":"10.1109\/TMI.2020.3001036"},{"doi-asserted-by":"publisher","key":"e_1_3_2_2_7_1","DOI":"10.1109\/TMI.2018.2806309"},{"doi-asserted-by":"publisher","key":"e_1_3_2_2_8_1","DOI":"10.1109\/CVPR.2016.90"},{"doi-asserted-by":"publisher","key":"e_1_3_2_2_9_1","DOI":"10.1016\/j.ipm.2020.102352"},{"doi-asserted-by":"publisher","key":"e_1_3_2_2_10_1","DOI":"10.1109\/ICASSP40776.2020.9053405"},{"key":"e_1_3_2_2_11_1","volume-title":"MISS-Former: An effective medical image segmentation Transformer. arXiv preprint arXiv:2109.07162","author":"Huang Xiaohong","year":"2021","unstructured":"Xiaohong Huang , Zhifang Deng , Dandan Li , and Xueguang Yuan . 2021. MISS-Former: An effective medical image segmentation Transformer. arXiv preprint arXiv:2109.07162 ( 2021 ). Xiaohong Huang, Zhifang Deng, Dandan Li, and Xueguang Yuan. 2021. MISS-Former: An effective medical image segmentation Transformer. arXiv preprint arXiv:2109.07162 (2021)."},{"doi-asserted-by":"publisher","key":"e_1_3_2_2_12_1","DOI":"10.1016\/j.media.2017.03.006"},{"doi-asserted-by":"publisher","key":"e_1_3_2_2_13_1","DOI":"10.1109\/TIP.2015.2481326"},{"doi-asserted-by":"publisher","key":"e_1_3_2_2_14_1","DOI":"10.1109\/ICCV48922.2021.00986"},{"doi-asserted-by":"publisher","key":"e_1_3_2_2_15_1","DOI":"10.1007\/s10462-011-9220-3"},{"doi-asserted-by":"publisher","key":"e_1_3_2_2_16_1","DOI":"10.1016\/j.media.2015.06.009"},{"key":"e_1_3_2_2_17_1","volume-title":"Matthew Lee, Mattias Heinrich, Kazunari Misawa, Kensaku Mori, Steven McDonagh, Nils Y Hammerla","author":"Oktay Ozan","year":"2018","unstructured":"Ozan Oktay , Jo Schlemper , Loic Le Folgoc , Matthew Lee, Mattias Heinrich, Kazunari Misawa, Kensaku Mori, Steven McDonagh, Nils Y Hammerla , Bernhard Kainz , et al . 2018 . Attention u-net: Learning where to look for the pancreas. arXiv preprint arXiv:1804.03999 (2018). Ozan Oktay, Jo Schlemper, Loic Le Folgoc, Matthew Lee, Mattias Heinrich, Kazunari Misawa, Kensaku Mori, Steven McDonagh, Nils Y Hammerla, Bernhard Kainz, et al . 2018. Attention u-net: Learning where to look for the pancreas. arXiv preprint arXiv:1804.03999 (2018)."},{"doi-asserted-by":"publisher","key":"e_1_3_2_2_18_1","DOI":"10.1016\/j.media.2018.02.001"},{"key":"e_1_3_2_2_19_1","volume-title":"Conformer: Local Features Coupling Global Representations for Visual Recognition. arXiv preprint arXiv:2105.03889","author":"Peng Zhiliang","year":"2021","unstructured":"Zhiliang Peng , Wei Huang , Shanzhi Gu , Lingxi Xie , Yaowei Wang , Jianbin Jiao , and Qixiang Ye . 2021 . Conformer: Local Features Coupling Global Representations for Visual Recognition. arXiv preprint arXiv:2105.03889 (2021). Zhiliang Peng, Wei Huang, Shanzhi Gu, Lingxi Xie, Yaowei Wang, Jianbin Jiao, and Qixiang Ye. 2021. Conformer: Local Features Coupling Global Representations for Visual Recognition. arXiv preprint arXiv:2105.03889 (2021)."},{"doi-asserted-by":"publisher","key":"e_1_3_2_2_20_1","DOI":"10.1007\/978-3-030-87589-3_28"},{"doi-asserted-by":"publisher","key":"e_1_3_2_2_21_1","DOI":"10.1007\/978-3-319-24574-4_28"},{"key":"e_1_3_2_2_22_1","volume-title":"UCTransNet: Rethinking the skip connections in U-Net from a channel-wise perspective with transformer. arXiv preprint arXiv:2109.04335","author":"Wang Haonan","year":"2021","unstructured":"Haonan Wang , Peng Cao , Jiaqi Wang , and Osmar R Zaiane . 2021. UCTransNet: Rethinking the skip connections in U-Net from a channel-wise perspective with transformer. arXiv preprint arXiv:2109.04335 ( 2021 ). Haonan Wang, Peng Cao, Jiaqi Wang, and Osmar R Zaiane. 2021. UCTransNet: Rethinking the skip connections in U-Net from a channel-wise perspective with transformer. arXiv preprint arXiv:2109.04335 (2021)."},{"key":"e_1_3_2_2_23_1","volume-title":"Mixed Transformer U-Net For Medical Image Segmentation. arXiv preprint arXiv:2111.04734","author":"Wang Hongyi","year":"2021","unstructured":"Hongyi Wang , Shiao Xie , Lanfen Lin , Yutaro Iwamoto , Xian-Hua Han , Yen-Wei Chen , and Ruofeng Tong . 2021. Mixed Transformer U-Net For Medical Image Segmentation. arXiv preprint arXiv:2111.04734 ( 2021 ). Hongyi Wang, Shiao Xie, Lanfen Lin, Yutaro Iwamoto, Xian-Hua Han, Yen-Wei Chen, and Ruofeng Tong. 2021. Mixed Transformer U-Net For Medical Image Segmentation. arXiv preprint arXiv:2111.04734 (2021)."},{"doi-asserted-by":"publisher","key":"e_1_3_2_2_24_1","DOI":"10.1109\/ICCV48922.2021.00061"},{"doi-asserted-by":"publisher","key":"e_1_3_2_2_25_1","DOI":"10.1007\/978-3-030-00937-3_60"},{"doi-asserted-by":"publisher","key":"e_1_3_2_2_26_1","DOI":"10.1016\/j.media.2019.04.005"},{"doi-asserted-by":"publisher","key":"e_1_3_2_2_27_1","DOI":"10.1007\/978-3-030-01234-2_1"},{"key":"e_1_3_2_2_28_1","volume-title":"Cvt: Introducing convolutions to vision transformers. arXiv preprint arXiv:2103.15808","author":"Wu Haiping","year":"2021","unstructured":"Haiping Wu , Bin Xiao , Noel Codella , Mengchen Liu , Xiyang Dai , Lu Yuan , and Lei Zhang . 2021 . Cvt: Introducing convolutions to vision transformers. arXiv preprint arXiv:2103.15808 (2021). Haiping Wu, Bin Xiao, Noel Codella, Mengchen Liu, Xiyang Dai, Lu Yuan, and Lei Zhang. 2021. Cvt: Introducing convolutions to vision transformers. arXiv preprint arXiv:2103.15808 (2021)."},{"doi-asserted-by":"publisher","key":"e_1_3_2_2_29_1","DOI":"10.1109\/TMI.2019.2930679"},{"key":"e_1_3_2_2_30_1","volume-title":"CoTr: Efficiently Bridging CNN and Transformer for 3D Medical Image Segmentation. arXiv preprint arXiv:2103.03024","author":"Xie Yutong","year":"2021","unstructured":"Yutong Xie , Jianpeng Zhang , Chunhua Shen , and Yong Xia . 2021. CoTr: Efficiently Bridging CNN and Transformer for 3D Medical Image Segmentation. arXiv preprint arXiv:2103.03024 ( 2021 ). Yutong Xie, Jianpeng Zhang, Chunhua Shen, and Yong Xia. 2021. CoTr: Efficiently Bridging CNN and Transformer for 3D Medical Image Segmentation. arXiv preprint arXiv:2103.03024 (2021)."},{"key":"e_1_3_2_2_31_1","volume-title":"Levit-unet: Make faster encoders with transformer for medical image segmentation. arXiv preprint arXiv:2107.08623","author":"Xu Guoping","year":"2021","unstructured":"Guoping Xu , Xingrong Wu , Xuan Zhang , and Xinwei He . 2021 . Levit-unet: Make faster encoders with transformer for medical image segmentation. arXiv preprint arXiv:2107.08623 (2021). Guoping Xu, Xingrong Wu, Xuan Zhang, and Xinwei He. 2021. Levit-unet: Make faster encoders with transformer for medical image segmentation. arXiv preprint arXiv:2107.08623 (2021)."},{"key":"e_1_3_2_2_32_1","volume-title":"Epsanet: An efficient pyramid split attention block on convolutional neural network. arXiv preprint arXiv:2105.14447","author":"Zhang Hu","year":"2021","unstructured":"Hu Zhang , Keke Zu , Jian Lu , Yuru Zou , and Deyu Meng . 2021 . Epsanet: An efficient pyramid split attention block on convolutional neural network. arXiv preprint arXiv:2105.14447 (2021). Hu Zhang, Keke Zu, Jian Lu, Yuru Zou, and Deyu Meng. 2021. Epsanet: An efficient pyramid split attention block on convolutional neural network. arXiv preprint arXiv:2105.14447 (2021)."},{"doi-asserted-by":"publisher","key":"e_1_3_2_2_33_1","DOI":"10.1109\/TMI.2020.2975347"},{"doi-asserted-by":"publisher","key":"e_1_3_2_2_34_1","DOI":"10.1016\/j.eswa.2021.114566"},{"key":"e_1_3_2_2_35_1","volume-title":"Nima Tajbakhsh, and Jianming Liang.","author":"Zhou Zongwei","year":"2018","unstructured":"Zongwei Zhou , Md Mahfuzur Rahman Siddiquee , Nima Tajbakhsh, and Jianming Liang. 2018 . Unet++: A nested u-net architecture for medical image segmentation. In Deep Learning in Medical Image Analysis and Multimodal Learning for Clinical Decision Support. Springer , 3--11. Zongwei Zhou, Md Mahfuzur Rahman Siddiquee, Nima Tajbakhsh, and Jianming Liang. 2018. Unet++: A nested u-net architecture for medical image segmentation. In Deep Learning in Medical Image Analysis and Multimodal Learning for Clinical Decision Support. Springer, 3--11."}],"event":{"sponsor":["SIGMM ACM Special Interest Group on Multimedia"],"acronym":"ICMR '22","name":"ICMR '22: International Conference on Multimedia Retrieval","location":"Newark NJ USA"},"container-title":["Proceedings of the 2022 International Conference on Multimedia Retrieval"],"original-title":[],"link":[{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/3512527.3531420","content-type":"unspecified","content-version":"vor","intended-application":"text-mining"},{"URL":"https:\/\/dl.acm.org\/doi\/pdf\/10.1145\/3512527.3531420","content-type":"unspecified","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2025,6,17]],"date-time":"2025-06-17T19:30:13Z","timestamp":1750188613000},"score":1,"resource":{"primary":{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/3512527.3531420"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2022,6,27]]},"references-count":35,"alternative-id":["10.1145\/3512527.3531420","10.1145\/3512527"],"URL":"https:\/\/doi.org\/10.1145\/3512527.3531420","relation":{},"subject":[],"published":{"date-parts":[[2022,6,27]]},"assertion":[{"value":"2022-06-27","order":2,"name":"published","label":"Published","group":{"name":"publication_history","label":"Publication History"}}]}}