{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2026,7,17]],"date-time":"2026-07-17T00:36:27Z","timestamp":1784248587429,"version":"3.55.0"},"publisher-location":"New York, NY, USA","reference-count":51,"publisher":"ACM","funder":[{"DOI":"10.13039\/501100012226","name":"Fundamental Research Funds for the Central Universities","doi-asserted-by":"publisher","award":["232023D-19"],"award-info":[{"award-number":["232023D-19"]}],"id":[{"id":"10.13039\/501100012226","id-type":"DOI","asserted-by":"publisher"}]}],"content-domain":{"domain":["dl.acm.org"],"crossmark-restriction":true},"short-container-title":[],"published-print":{"date-parts":[[2025,10,27]]},"DOI":"10.1145\/3746027.3755210","type":"proceedings-article","created":{"date-parts":[[2025,10,25]],"date-time":"2025-10-25T07:26:51Z","timestamp":1761377211000},"page":"1500-1509","update-policy":"https:\/\/doi.org\/10.1145\/crossmark-policy","source":"Crossref","is-referenced-by-count":2,"title":["Bridging the Unseen Gap: Label-Enhanced Information Bottleneck Distillation for Multimodal Named Entity Recognition"],"prefix":"10.1145","author":[{"ORCID":"https:\/\/orcid.org\/0000-0002-2083-4307","authenticated-orcid":false,"given":"Bo","family":"Xu","sequence":"first","affiliation":[{"name":"School of Computer Science and Technology, Donghua University, Shanghai, China"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"ORCID":"https:\/\/orcid.org\/0009-0004-8007-8197","authenticated-orcid":false,"given":"Jie","family":"Wei","sequence":"additional","affiliation":[{"name":"School of Computer Science and Technology, Donghua University, Shanghai, China"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"ORCID":"https:\/\/orcid.org\/0000-0002-5409-9347","authenticated-orcid":false,"given":"Hongya","family":"Wang","sequence":"additional","affiliation":[{"name":"School of Computer Science and Technology, Donghua University, Shanghai, China"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"ORCID":"https:\/\/orcid.org\/0000-0002-8617-8939","authenticated-orcid":false,"given":"Ming","family":"Du","sequence":"additional","affiliation":[{"name":"School of Computer Science and Technology, Donghua University, Shanghai, China"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"ORCID":"https:\/\/orcid.org\/0009-0009-0436-1456","authenticated-orcid":false,"given":"Hui","family":"Song","sequence":"additional","affiliation":[{"name":"School of Computer Science and Technology, Donghua University, Shanghai, China"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"ORCID":"https:\/\/orcid.org\/0000-0001-8403-9591","authenticated-orcid":false,"given":"Yanghua","family":"Xiao","sequence":"additional","affiliation":[{"name":"Shanghai Key Laboratory of Data Science, College of Computer Science and Artificial Intelligence, Fudan University, Shanghai, China"}],"role":[{"vocabulary":"crossref","role":"author"}]}],"member":"320","published-online":{"date-parts":[[2025,10,27]]},"reference":[{"key":"e_1_3_2_1_1_1","unstructured":"Alexander A Alemi Ian Fischer Joshua V Dillon and Kevin Murphy. 2017. Deep variational information bottleneck. In ICLR."},{"key":"e_1_3_2_1_2_1","first-page":"531","article-title":"Mutual information neural estimation","author":"Belghazi Mohamed Ishmael","year":"2018","unstructured":"Mohamed Ishmael Belghazi, Aristide Baratin, Sai Rajeshwar, Sherjil Ozair, Yoshua Bengio, Aaron Courville, and Devon Hjelm. 2018. Mutual information neural estimation. In ICML. PMLR, 531-540.","journal-title":"ICML. PMLR"},{"key":"e_1_3_2_1_3_1","first-page":"1607","volume-title":"Good Visual Guidance Make A Better Extractor: Hierarchical Visual Prefix for Multimodal Entity and Relation Extraction. In Findings of the Association for Computational Linguistics: NAACL","author":"Chen Xiang","year":"2022","unstructured":"Xiang Chen, Ningyu Zhang, Lei Li, Yunzhi Yao, Shumin Deng, Chuanqi Tan, Fei Huang, Luo Si, and Huajun Chen. 2022a. Good Visual Guidance Make A Better Extractor: Hierarchical Visual Prefix for Multimodal Entity and Relation Extraction. In Findings of the Association for Computational Linguistics: NAACL 2022, Marine Carpuat, Marie-Catherine de Marneffe, and Ivan Vladimir Meza Ruiz (Eds.). Association for Computational Linguistics, Seattle, United States, 1607-1618."},{"key":"e_1_3_2_1_4_1","volume-title":"Good visual guidance makes a better extractor: Hierarchical visual prefix for multimodal entity and relation extraction. arXiv preprint arXiv:2205.03521","author":"Chen Xiang","year":"2022","unstructured":"Xiang Chen, Ningyu Zhang, Lei Li, Yunzhi Yao, Shumin Deng, Chuanqi Tan, Fei Huang, Luo Si, and Huajun Chen. 2022b. Good visual guidance makes a better extractor: Hierarchical visual prefix for multimodal entity and relation extraction. arXiv preprint arXiv:2205.03521 (2022)."},{"key":"e_1_3_2_1_5_1","first-page":"1779","article-title":"Club: A contrastive log-ratio upper bound of mutual information","author":"Cheng Pengyu","year":"2020","unstructured":"Pengyu Cheng, Weituo Hao, Shuyang Dai, Jiachang Liu, Zhe Gan, and Lawrence Carin. 2020. Club: A contrastive log-ratio upper bound of mutual information. In ICML. PMLR, 1779-1788.","journal-title":"ICML. PMLR"},{"key":"e_1_3_2_1_6_1","doi-asserted-by":"publisher","DOI":"10.1109\/TASLP.2023.3345146"},{"key":"e_1_3_2_1_7_1","volume-title":"Bert: Pre-training of deep bidirectional transformers for language understanding. arXiv preprint arXiv:1810.04805","author":"Devlin Jacob","year":"2018","unstructured":"Jacob Devlin. 2018. Bert: Pre-training of deep bidirectional transformers for language understanding. arXiv preprint arXiv:1810.04805 (2018)."},{"key":"e_1_3_2_1_8_1","first-page":"18674","article-title":"Learning optimal representations with the decodable information bottleneck","volume":"33","author":"Dubois Yann","year":"2020","unstructured":"Yann Dubois, Douwe Kiela, David J Schwab, and Ramakrishna Vedantam. 2020. Learning optimal representations with the decodable information bottleneck. Advances in Neural Information Processing Systems, Vol. 33 (2020), 18674-18690.","journal-title":"Advances in Neural Information Processing Systems"},{"key":"e_1_3_2_1_9_1","first-page":"4827","volume-title":"Robust Backed-off Estimation of Out-of-Vocabulary Embeddings. In Findings of the Association for Computational Linguistics: EMNLP","author":"Fukuda Nobukazu","year":"2020","unstructured":"Nobukazu Fukuda, Naoki Yoshinaga, and Masaru Kitsuregawa. 2020. Robust Backed-off Estimation of Out-of-Vocabulary Embeddings. In Findings of the Association for Computational Linguistics: EMNLP 2020, Trevor Cohn, Yulan He, and Yang Liu (Eds.). Association for Computational Linguistics, Online, 4827-4838."},{"key":"e_1_3_2_1_10_1","doi-asserted-by":"publisher","DOI":"10.1109\/ISI58743.2023.10297238"},{"key":"e_1_3_2_1_11_1","doi-asserted-by":"publisher","DOI":"10.1609\/aaai.v35i13.17370"},{"key":"e_1_3_2_1_12_1","volume-title":"The variational bandwidth bottleneck: Stochastic evaluation on an information budget. arXiv preprint arXiv:2004.11935","author":"Goyal Anirudh","year":"2020","unstructured":"Anirudh Goyal, Yoshua Bengio, Matthew Botvinick, and Sergey Levine. 2020. The variational bandwidth bottleneck: Stochastic evaluation on an information budget. arXiv preprint arXiv:2004.11935 (2020)."},{"key":"e_1_3_2_1_13_1","doi-asserted-by":"publisher","DOI":"10.1145\/3583780.3614967"},{"key":"e_1_3_2_1_14_1","doi-asserted-by":"publisher","DOI":"10.1609\/aaai.v34i05.6299"},{"key":"e_1_3_2_1_15_1","volume-title":"Stochastic neighbor embedding. Advances in neural information processing systems","author":"Hinton Geoffrey E","year":"2002","unstructured":"Geoffrey E Hinton and Sam Roweis. 2002. Stochastic neighbor embedding. Advances in neural information processing systems, Vol. 15 (2002)."},{"key":"e_1_3_2_1_16_1","volume-title":"Bidirectional LSTM-CRF Models for Sequence Tagging. arXiv preprint arXiv:1508.01991","author":"Huang Zhiheng","year":"2015","unstructured":"Zhiheng Huang. 2015. Bidirectional LSTM-CRF Models for Sequence Tagging. arXiv preprint arXiv:1508.01991 (2015)."},{"key":"e_1_3_2_1_17_1","doi-asserted-by":"publisher","DOI":"10.1609\/aaai.v37i7.25971"},{"key":"e_1_3_2_1_18_1","doi-asserted-by":"publisher","DOI":"10.1145\/3503161.3548427"},{"key":"e_1_3_2_1_19_1","volume-title":"Auto-encoding variational bayes. arXiv preprint arXiv:1312.6114","author":"Kingma Diederik P","year":"2013","unstructured":"Diederik P Kingma. 2013. Auto-encoding variational bayes. arXiv preprint arXiv:1312.6114 (2013)."},{"key":"e_1_3_2_1_20_1","volume-title":"Icml","volume":"1","author":"Lafferty John","year":"2001","unstructured":"John Lafferty, Andrew McCallum, Fernando Pereira, et al., 2001. Conditional random fields: Probabilistic models for segmenting and labeling sequence data. In Icml, Vol. 1. Williamstown, MA, 3."},{"key":"e_1_3_2_1_21_1","volume-title":"Proceedings of the 40th International Conference on Machine Learning (Proceedings of Machine Learning Research","volume":"19742","author":"Li Junnan","year":"2023","unstructured":"Junnan Li, Dongxu Li, Silvio Savarese, and Steven Hoi. 2023. BLIP-2: Bootstrapping Language-Image Pre-training with Frozen Image Encoders and Large Language Models. In Proceedings of the 40th International Conference on Machine Learning (Proceedings of Machine Learning Research, Vol. 202), Andreas Krause, Emma Brunskill, Kyunghyun Cho, Barbara Engelhardt, Sivan Sabato, and Jonathan Scarlett (Eds.). PMLR, 19730-19742."},{"key":"e_1_3_2_1_22_1","volume-title":"Proceedings of the 61st Annual Meeting of the Association for Computational Linguistics (Volume 1: Long Papers), Anna Rogers, Jordan Boyd-Graber","author":"Liang Ziran","unstructured":"Ziran Liang, Yuyin Lu, HeGang Chen, and Yanghui Rao. 2023. Graph-based Relation Mining for Context-free Out-of-vocabulary Word Embedding Learning. In Proceedings of the 61st Annual Meeting of the Association for Computational Linguistics (Volume 1: Long Papers), Anna Rogers, Jordan Boyd-Graber, and Naoaki Okazaki (Eds.). Association for Computational Linguistics, Toronto, Canada, 14133-14149."},{"key":"e_1_3_2_1_23_1","doi-asserted-by":"publisher","DOI":"10.18653\/v1\/P18-1185"},{"key":"e_1_3_2_1_24_1","doi-asserted-by":"publisher","DOI":"10.1609\/aaai.v34i04.5950"},{"key":"e_1_3_2_1_25_1","doi-asserted-by":"publisher","DOI":"10.18653\/v1\/N18-1078"},{"key":"e_1_3_2_1_26_1","doi-asserted-by":"publisher","DOI":"10.1103\/PhysRevE.52.2318"},{"key":"e_1_3_2_1_27_1","volume-title":"SCANNER: Knowledge-Enhanced Approach for Robust Multi-modal Named Entity Recognition of Unseen Entities. arXiv preprint arXiv:2404.01914","author":"Ok Hyunjong","year":"2024","unstructured":"Hyunjong Ok, Taeho Kil, Sukmin Seo, and Jaeho Lee. 2024. SCANNER: Knowledge-Enhanced Approach for Robust Multi-modal Named Entity Recognition of Unseen Entities. arXiv preprint arXiv:2404.01914 (2024)."},{"key":"e_1_3_2_1_28_1","volume-title":"Representation learning with contrastive predictive coding. arXiv preprint arXiv:1807.03748","author":"van den Oord Aaron","year":"2018","unstructured":"Aaron van den Oord, Yazhe Li, and Oriol Vinyals. 2018. Representation learning with contrastive predictive coding. arXiv preprint arXiv:1807.03748 (2018)."},{"key":"e_1_3_2_1_29_1","volume-title":"International Conference on Machine Learning. PMLR, 5171-5180","author":"Poole Ben","year":"2019","unstructured":"Ben Poole, Sherjil Ozair, Aaron Van Den Oord, Alex Alemi, and George Tucker. 2019. On variational bounds of mutual information. In International Conference on Machine Learning. PMLR, 5171-5180."},{"key":"e_1_3_2_1_30_1","volume-title":"International conference on machine learning. PMLR, 8748-8763","author":"Radford Alec","year":"2021","unstructured":"Alec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh, Gabriel Goh, Sandhini Agarwal, Girish Sastry, Amanda Askell, Pamela Mishkin, Jack Clark, et al., 2021. Learning transferable visual models from natural language supervision. In International conference on machine learning. PMLR, 8748-8763."},{"key":"e_1_3_2_1_31_1","volume-title":"Proceedings of the 2011 conference on empirical methods in natural language processing. 1524-1534","author":"Ritter Alan","year":"2011","unstructured":"Alan Ritter, Sam Clark, Oren Etzioni, et al., 2011. Named entity recognition in tweets: an experimental study. In Proceedings of the 2011 conference on empirical methods in natural language processing. 1524-1534."},{"key":"e_1_3_2_1_32_1","doi-asserted-by":"publisher","DOI":"10.18653\/v1\/N19-1048"},{"key":"e_1_3_2_1_33_1","volume-title":"Neural architectures for named entity recognition. CoRR abs\/1603.01360","author":"Subramanian S","year":"2016","unstructured":"S Subramanian, K Kawakami, and C Dyer. 2016. Neural architectures for named entity recognition. CoRR abs\/1603.01360 (2016)."},{"key":"e_1_3_2_1_34_1","volume-title":"The information bottleneck method. arXiv preprint physics\/0004057","author":"Tishby Naftali","year":"2000","unstructured":"Naftali Tishby, Fernando C Pereira, and William Bialek. 2000. The information bottleneck method. arXiv preprint physics\/0004057 (2000)."},{"key":"e_1_3_2_1_35_1","first-page":"1","article-title":"Deep learning and the information bottleneck principle. In 2015 ieee information theory workshop (itw)","author":"Tishby Naftali","year":"2015","unstructured":"Naftali Tishby and Noga Zaslavsky. 2015. Deep learning and the information bottleneck principle. In 2015 ieee information theory workshop (itw). IEEE, 1-5.","journal-title":"IEEE"},{"key":"e_1_3_2_1_36_1","volume-title":"Tjong Kim Sang and Jorn Veenstra","author":"Erik","year":"1999","unstructured":"Erik F. Tjong Kim Sang and Jorn Veenstra. 1999. Representing Text Chunks. In Ninth Conference of the European Chapter of the Association for Computational Linguistics, Henry S. Thompson and Alex Lascarides (Eds.). Association for Computational Linguistics, Bergen, Norway, 173-179."},{"key":"e_1_3_2_1_37_1","volume-title":"Proceedings of the 39th ACM\/SIGAPP Symposium on Applied Computing. 1479-1486","author":"Anu Mary Chacko Akhila","year":"2024","unstructured":"Akhila VH and Anu Mary Chacko. 2024. Cooperative Embedding-A Novel Approach to Tackle the Out-Of-Vocabulary Dilemma in Bot Classification. In Proceedings of the 39th ACM\/SIGAPP Symposium on Applied Computing. 1479-1486."},{"key":"e_1_3_2_1_38_1","doi-asserted-by":"publisher","DOI":"10.1609\/aaai.v35i11.17210"},{"key":"e_1_3_2_1_39_1","doi-asserted-by":"crossref","unstructured":"Hairong Wang Tong Wang Chong Sun et al. 2023b. Multi-Scale Visual Semantic Enhanced for Multi-Modal Ner. Tong and Sun Chong Multi-Scale Visual Semantic Enhanced for Multi-Modal Ner (2023).","DOI":"10.2139\/ssrn.4656122"},{"key":"e_1_3_2_1_40_1","volume-title":"Clglf: Confidence Learning Guides Label Fusion for Multimodal Named Entity Recognition Method. Available at SSRN 4813568","author":"Wang Hairong","year":"2024","unstructured":"Hairong Wang, Tong Wang, Yiyan Wang, and Fangping Chen. 2024. Clglf: Confidence Learning Guides Label Fusion for Multimodal Named Entity Recognition Method. Available at SSRN 4813568 (2024)."},{"key":"e_1_3_2_1_41_1","volume-title":"ITA: Image-text alignments for multi-modal named entity recognition. arXiv preprint arXiv:2112.06482","author":"Wang Xinyu","year":"2021","unstructured":"Xinyu Wang, Min Gui, Yong Jiang, Zixia Jia, Nguyen Bach, Tao Wang, Zhongqiang Huang, Fei Huang, and Kewei Tu. 2021. ITA: Image-text alignments for multi-modal named entity recognition. arXiv preprint arXiv:2112.06482 (2021)."},{"key":"e_1_3_2_1_42_1","doi-asserted-by":"publisher","DOI":"10.1109\/ICME52920.2022.9859972"},{"key":"e_1_3_2_1_43_1","doi-asserted-by":"publisher","DOI":"10.18653\/v1\/2023.conll-1.27"},{"key":"e_1_3_2_1_44_1","first-page":"20437","article-title":"Graph information bottleneck","volume":"33","author":"Wu Tailin","year":"2020","unstructured":"Tailin Wu, Hongyu Ren, Pan Li, and Jure Leskovec. 2020. Graph information bottleneck. NeurIPs, Vol. 33 (2020), 20437-20448.","journal-title":"NeurIPs"},{"key":"e_1_3_2_1_45_1","doi-asserted-by":"publisher","DOI":"10.1145\/3488560.3498475"},{"key":"e_1_3_2_1_46_1","volume-title":"Proceedings of the 58th Annual Meeting of the Association for Computational Linguistics","author":"Yu Jianfei","unstructured":"Jianfei Yu, Jing Jiang, Li Yang, and Rui Xia. 2020. Improving Multimodal Named Entity Recognition via Entity Span Detection with Unified Multimodal Transformer. In Proceedings of the 58th Annual Meeting of the Association for Computational Linguistics, Dan Jurafsky, Joyce Chai, Natalie Schluter, and Joel Tetreault (Eds.). Association for Computational Linguistics, Online, 3342-3352."},{"key":"e_1_3_2_1_47_1","first-page":"1650","volume-title":"IEEE TPAMI","volume":"46","author":"Yu Junchi","year":"2021","unstructured":"Junchi Yu, Tingyang Xu, Yu Rong, Yatao Bian, Junzhou Huang, and Ran He. 2021. Recognizing predictive substructures with subgraph information bottleneck. IEEE TPAMI, Vol. 46, 3 (2021), 1650-1663."},{"key":"e_1_3_2_1_48_1","doi-asserted-by":"publisher","DOI":"10.1609\/aaai.v35i16.17687"},{"key":"e_1_3_2_1_49_1","doi-asserted-by":"publisher","DOI":"10.1609\/aaai.v32i1.11962"},{"key":"e_1_3_2_1_50_1","doi-asserted-by":"publisher","DOI":"10.1145\/3539597.3570485"},{"key":"e_1_3_2_1_51_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR52688.2022.01631"}],"event":{"name":"MM '25: The 33rd ACM International Conference on Multimedia","location":"Dublin Ireland","acronym":"MM '25","sponsor":["SIGMM ACM Special Interest Group on Multimedia"]},"container-title":["Proceedings of the 33rd ACM International Conference on Multimedia"],"original-title":[],"link":[{"URL":"https:\/\/dl.acm.org\/doi\/pdf\/10.1145\/3746027.3755210","content-type":"unspecified","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2025,12,10]],"date-time":"2025-12-10T04:09:06Z","timestamp":1765339746000},"score":1,"resource":{"primary":{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/3746027.3755210"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2025,10,27]]},"references-count":51,"alternative-id":["10.1145\/3746027.3755210","10.1145\/3746027"],"URL":"https:\/\/doi.org\/10.1145\/3746027.3755210","relation":{},"subject":[],"published":{"date-parts":[[2025,10,27]]},"assertion":[{"value":"2025-10-27","order":3,"name":"published","label":"Published","group":{"name":"publication_history","label":"Publication History"}}]}}