{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2026,6,16]],"date-time":"2026-06-16T05:27:51Z","timestamp":1781587671383,"version":"3.54.5"},"reference-count":63,"publisher":"Association for Computing Machinery (ACM)","issue":"12","content-domain":{"domain":["dl.acm.org"],"crossmark-restriction":true},"short-container-title":["ACM Trans. Asian Low-Resour. Lang. Inf. Process."],"published-print":{"date-parts":[[2025,12,31]]},"abstract":"<jats:p>Metaphors play a crucial role in human communication, yet their comprehension remains a significant challenge for natural language processing (NLP) due to the cognitive complexity involved. According to Conceptual Metaphor Theory (CMT), metaphors map a target domain onto a source domain, and understanding this mapping is essential for grasping the nature of metaphors. Existing NLP research has focused on tasks like metaphor detection and sentiment analysis. However, there has been limited attention to identifying mappings between source and target domains. Moreover, non-English multimodal metaphor resources remain largely neglected in the literature, hindering a deeper understanding of the key elements involved in metaphor interpretation. To address this gap, we developed a Chinese multimodal metaphor advertisement dataset (namely CM3D) that includes annotations of specific target and source domains. This dataset aims at fostering further research into metaphor comprehension, particularly in non-English languages. Furthermore, we propose a Chain-of-Thought (CoT) Prompting-based Metaphor Mapping Identification Model (CPMMIM), which simulates the human cognitive process for identifying these mappings. Drawing inspiration from CoT reasoning and Bi-Level Optimization (BLO), we treat the task as a hierarchical identification problem, enabling more accurate and interpretable metaphor mapping. Our experimental results demonstrate the effectiveness of CPMMIM, highlighting its potential for advancing metaphor comprehension in NLP. Our dataset and code are both publicly available to encourage further advancements in this field.<\/jats:p>","DOI":"10.1145\/3773989","type":"journal-article","created":{"date-parts":[[2025,10,31]],"date-time":"2025-10-31T06:19:56Z","timestamp":1761891596000},"page":"1-25","update-policy":"https:\/\/doi.org\/10.1145\/crossmark-policy","source":"Crossref","is-referenced-by-count":6,"title":["Towards Multimodal Metaphor Understanding: A Chinese Dataset and Model for Metaphor Mapping Identification"],"prefix":"10.1145","volume":"24","author":[{"ORCID":"https:\/\/orcid.org\/0000-0002-7683-5560","authenticated-orcid":false,"given":"Dongyu","family":"Zhang","sequence":"first","affiliation":[{"name":"School of Foreign Languages, Dalian University of Technology","place":["Dalian, China"]}],"role":[{"vocabulary":"crossref","role":"author"}]},{"ORCID":"https:\/\/orcid.org\/0009-0005-4830-9968","authenticated-orcid":false,"given":"Shengcheng","family":"Yin","sequence":"additional","affiliation":[{"name":"School of Software, Dalian University of Technology","place":["Dalian, China"]}],"role":[{"vocabulary":"crossref","role":"author"}]},{"ORCID":"https:\/\/orcid.org\/0009-0002-1045-1318","authenticated-orcid":false,"given":"Jingwei","family":"Yu","sequence":"additional","affiliation":[{"name":"School of Software, Dalian University of Technology","place":["Dalian, China"]}],"role":[{"vocabulary":"crossref","role":"author"}]},{"ORCID":"https:\/\/orcid.org\/0009-0005-0827-0720","authenticated-orcid":false,"given":"Zhiyao","family":"Wu","sequence":"additional","affiliation":[{"name":"Faculty of Business Administration, University of Macau","place":["Macau, China"]}],"role":[{"vocabulary":"crossref","role":"author"}]},{"ORCID":"https:\/\/orcid.org\/0000-0001-7751-485X","authenticated-orcid":false,"given":"Zhen","family":"Li","sequence":"additional","affiliation":[{"name":"Faculty of Business and Commerce, Kansai University","place":["Suita, Japan"]}],"role":[{"vocabulary":"crossref","role":"author"}]},{"ORCID":"https:\/\/orcid.org\/0000-0001-8860-7701","authenticated-orcid":false,"given":"Chengpei","family":"Xu","sequence":"additional","affiliation":[{"name":"School of Minerals and Energy Resources Engineering, University of New South Wales","place":["Sydney, Australia"]}],"role":[{"vocabulary":"crossref","role":"author"}]},{"ORCID":"https:\/\/orcid.org\/0000-0003-4743-1994","authenticated-orcid":false,"given":"Xiaoxia","family":"Wang","sequence":"additional","affiliation":[{"name":"Centre for Educational Innovation and Quality, RMIT University","place":["Melbourne, Australia"]}],"role":[{"vocabulary":"crossref","role":"author"}]},{"ORCID":"https:\/\/orcid.org\/0000-0002-8324-1859","authenticated-orcid":false,"given":"Feng","family":"Xia","sequence":"additional","affiliation":[{"name":"School of Computing Technologies, RMIT University","place":["Melbourne, Australia"]}],"role":[{"vocabulary":"crossref","role":"author"}]}],"member":"320","published-online":{"date-parts":[[2025,12,10]]},"reference":[{"key":"e_1_3_2_2_2","first-page":"9","article-title":"A method for linguistic metaphor identification from MIP to MIPVU preface","author":"Steen Gerard","year":"2010","unstructured":"Gerard Steen, Lettie Dorst, Berenike Herrmann, Anna Kaal, Tina Krennmayr, and Trijntje Pasma. 2010. A method for linguistic metaphor identification from MIP to MIPVU preface. Method for Linguistic Metaphor Identification: From MIP To MIPVU Vol. 14. John Benjamins Publishing Company, Amsterdam\/Philadelphia, 9\u201320.","journal-title":"Method for Linguistic Metaphor Identification: From MIP To MIPVU"},{"key":"e_1_3_2_3_2","volume-title":"Metaphors We Live By","author":"Lakoff G.","year":"1980","unstructured":"G. Lakoff and M. Johnson. 1980. Metaphors We Live By. University of Chicago Press."},{"key":"e_1_3_2_4_2","doi-asserted-by":"publisher","DOI":"10.1145\/3664647.3681060"},{"key":"e_1_3_2_5_2","doi-asserted-by":"publisher","DOI":"10.1145\/3477495.3532019"},{"key":"e_1_3_2_6_2","doi-asserted-by":"publisher","DOI":"10.1109\/ACCESS.2013.2281156"},{"key":"e_1_3_2_7_2","doi-asserted-by":"publisher","DOI":"10.1515\/9783110215366"},{"key":"e_1_3_2_8_2","doi-asserted-by":"publisher","DOI":"10.18653\/v1\/2020.figlang-1.26"},{"key":"e_1_3_2_9_2","doi-asserted-by":"publisher","DOI":"10.18653\/v1\/2020.figlang-1.34"},{"key":"e_1_3_2_10_2","doi-asserted-by":"publisher","DOI":"10.18653\/v1\/2021.naacl-main.141"},{"key":"e_1_3_2_11_2","doi-asserted-by":"publisher","DOI":"10.1016\/j.jneuroling.2012.10.004"},{"key":"e_1_3_2_12_2","doi-asserted-by":"publisher","DOI":"10.18653\/v1\/S16-2003"},{"key":"e_1_3_2_13_2","doi-asserted-by":"publisher","DOI":"10.18653\/v1\/W18-0912"},{"key":"e_1_3_2_14_2","doi-asserted-by":"publisher","DOI":"10.18653\/v1\/2023.acl-long.58"},{"key":"e_1_3_2_15_2","doi-asserted-by":"publisher","DOI":"10.1162\/COLI_a_00275"},{"key":"e_1_3_2_16_2","doi-asserted-by":"publisher","DOI":"10.18653\/v1\/2021.acl-long.249"},{"key":"e_1_3_2_17_2","first-page":"4221","volume-title":"Proceedings of the 10th International Conference on Language Resources and Evaluation.","author":"Mohler Michael","year":"2016","unstructured":"Michael Mohler, Mary Brunson, Bryan Rink, and Marc Tomlinson. 2016. Introducing the lcc metaphor datasets. In Proceedings of the 10th International Conference on Language Resources and Evaluation.4221\u20134227."},{"key":"e_1_3_2_18_2","doi-asserted-by":"publisher","DOI":"10.3115\/v1\/W15-1405"},{"key":"e_1_3_2_19_2","doi-asserted-by":"publisher","DOI":"10.1016\/j.inffus.2022.06.002"},{"key":"e_1_3_2_20_2","doi-asserted-by":"publisher","DOI":"10.1609\/aaai.v36i10.21313"},{"key":"e_1_3_2_21_2","doi-asserted-by":"publisher","DOI":"10.1007\/s11063-024-11609-w"},{"key":"e_1_3_2_22_2","doi-asserted-by":"publisher","DOI":"10.18653\/v1\/2020.figlang-1.3"},{"key":"e_1_3_2_23_2","doi-asserted-by":"publisher","DOI":"10.1145\/3551349.3559555"},{"key":"e_1_3_2_24_2","doi-asserted-by":"publisher","DOI":"10.18653\/v1\/2021.findings-acl.366"},{"key":"e_1_3_2_25_2","doi-asserted-by":"publisher","DOI":"10.18653\/v1\/2022.naacl-main.330"},{"key":"e_1_3_2_26_2","doi-asserted-by":"publisher","DOI":"10.1109\/URTC60662.2023.10534945"},{"key":"e_1_3_2_27_2","volume-title":"Proceedings of the 36th International Conference on Neural Information Processing Systems.","author":"Wei Jason","year":"2024","unstructured":"Jason Wei, Xuezhi Wang, Dale Schuurmans, Maarten Bosma, Brian Ichter, Fei Xia, Ed H. Chi, Quoc V. Le, and Denny Zhou. 2024. Chain-of-thought prompting elicits reasoning in large language models. In Proceedings of the 36th International Conference on Neural Information Processing Systems.Curran Associates Inc., Red Hook, NY, USA, Article 1800, 14 pages."},{"key":"e_1_3_2_28_2","doi-asserted-by":"publisher","DOI":"10.1609\/aaai.v38i17.29884"},{"key":"e_1_3_2_29_2","volume-title":"Proceedings of the AAAI Conference on Artificial Intelligence","author":"Cohn Clayton","year":"2024","unstructured":"Clayton Cohn, Nicole M. Hutchins, Tuan Le, and Gautam Biswas. 2024. A chain-of-thought prompting approach with LLMs for evaluating students\u2019 formative assessment responses in science. In Proceedings of the AAAI Conference on Artificial Intelligence. Retrieved from https:\/\/api.semanticscholar.org\/CorpusID:268553761"},{"key":"e_1_3_2_30_2","doi-asserted-by":"publisher","DOI":"10.1109\/IALP61005.2023.10337264"},{"key":"e_1_3_2_31_2","unstructured":"Maxwell I. Nye Anders Johan Andreassen Guy Gur-Ari Henryk Michalewski Jacob Austin David Bieber David Dohan Aitor Lewkowycz Maarten Bosma David Luan Charles Sutton and Augustus Odena. 2021. Show your work: Scratchpads for intermediate computation with language models. Retrieved from arXiv:2112.00114.https:\/\/arxiv.org\/abs\/2112.00114"},{"key":"e_1_3_2_32_2","first-page":"24824","article-title":"Chain-of-thought prompting elicits reasoning in large language models","volume":"35","author":"Wei Jason","year":"2022","unstructured":"Jason Wei, Xuezhi Wang, Dale Schuurmans, Maarten Bosma, Fei Xia, Ed Chi, Quoc V. Le, Denny Zhou, et\u00a0al. 2022. Chain-of-thought prompting elicits reasoning in large language models. In Proceedings of the 36th International Conference on Neural Information Processing Systems 35 (2022), 24824\u201324837.","journal-title":"Proceedings of the 36th International Conference on Neural Information Processing Systems"},{"key":"e_1_3_2_33_2","first-page":"22199","volume-title":"Proceedings of the Advances in Neural Information Processing Systems","author":"Kojima Takeshi","year":"2022","unstructured":"Takeshi Kojima, Shixiang (Shane) Gu, Machel Reid, Yutaka Matsuo, and Yusuke Iwasawa. 2022. Large language models are zero-shot reasoners. In Proceedings of the Advances in Neural Information Processing Systems. 22199\u201322213."},{"key":"e_1_3_2_34_2","doi-asserted-by":"publisher","DOI":"10.1609\/aaai.v38i16.29794"},{"key":"e_1_3_2_35_2","doi-asserted-by":"publisher","DOI":"10.18653\/v1\/2023.acl-short.101"},{"key":"e_1_3_2_36_2","doi-asserted-by":"publisher","DOI":"10.1016\/j.engappai.2024.107907"},{"key":"e_1_3_2_37_2","unstructured":"Zhuosheng Zhang Aston Zhang Mu Li Hai Zhao George Karypis and Alex Smola. 2023. Multimodal chain-of-thought reasoning in language models. arXiv:2302.00923. Retrieved from https:\/\/arxiv.org\/abs\/2302.00923"},{"key":"e_1_3_2_38_2","first-page":"6059","volume-title":"Proceedings of the 2024 Joint International Conference on Computational Linguistics, Language Resources and Evaluation","author":"Zheng Guangmin","year":"2024","unstructured":"Guangmin Zheng, Jin Wang, Xiaobing Zhou, and Xuejie Zhang. 2024. Enhancing semantics in multimodal chain of thought via soft negative sampling. In Proceedings of the 2024 Joint International Conference on Computational Linguistics, Language Resources and Evaluation. Nicoletta Calzolari, Min-Yen Kan, Veronique Hoste, Alessandro Lenci, Sakriani Sakti, and Nianwen Xue (Eds.), ELRA and ICCL, Torino, Italia, 6059\u20136076. Retrieved from https:\/\/aclanthology.org\/2024.lrec-main.537"},{"key":"e_1_3_2_39_2","doi-asserted-by":"publisher","DOI":"10.1609\/aaai.v38i16.29776"},{"key":"e_1_3_2_40_2","doi-asserted-by":"publisher","DOI":"10.1073\/pnas.2218523120"},{"key":"e_1_3_2_41_2","unstructured":"Ishita Dasgupta Andrew Kyle Lampinen Stephanie C. Y. Chan Antonia Creswell Dharshan Kumaran James L. McClelland and Felix Hill. 2022. Language models show human-like content effects on reasoning. arXiv:2207.07051. Retrieved from https:\/\/arxiv.org\/abs\/2207.07051 Retrieved from https:\/\/api.semanticscholar.org\/CorpusID:250526626"},{"key":"e_1_3_2_42_2","doi-asserted-by":"publisher","DOI":"10.1073\/pnas.1407479111"},{"key":"e_1_3_2_43_2","unstructured":"Ben Prystawski Paul H. Thibodeau Christopher Potts and Noah D. Goodman. 2023. Psychologically-informed chain-of-thought prompts for metaphor understanding in large language models. In Proceedings of the Annual Meeting of the Cognitive Science Society. 2311\u20132317. Retrieved from https:\/\/escholarship.org\/uc\/item\/2q01t47h"},{"key":"e_1_3_2_44_2","doi-asserted-by":"publisher","DOI":"10.1109\/TPAMI.2021.3132674"},{"key":"e_1_3_2_45_2","doi-asserted-by":"publisher","DOI":"10.1007\/978-3-030-52119-6_20"},{"key":"e_1_3_2_46_2","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR52688.2022.00571"},{"key":"e_1_3_2_47_2","doi-asserted-by":"publisher","DOI":"10.1109\/TIP.2022.3140607"},{"key":"e_1_3_2_48_2","article-title":"Visual and multimodal metaphor in advertising: Cultural perspectives.","author":"Forceville Charles","year":"2017","unstructured":"Charles Forceville. 2017. Visual and multimodal metaphor in advertising: Cultural perspectives. Styles of Communication 9, 2 (2017).","journal-title":"Styles of Communication"},{"key":"e_1_3_2_49_2","doi-asserted-by":"publisher","DOI":"10.18653\/v1\/2023.findings-emnlp.409"},{"key":"e_1_3_2_50_2","doi-asserted-by":"publisher","DOI":"10.1177\/001316446002000104"},{"key":"e_1_3_2_51_2","doi-asserted-by":"publisher","DOI":"10.3115\/v1\/D14-1162"},{"key":"e_1_3_2_52_2","article-title":"UMAP: Uniform manifold approximation and projection for dimension reduction","author":"McInnes Leland","year":"2020","unstructured":"Leland McInnes, John Healy, and James Melville. 2020. UMAP: Uniform manifold approximation and projection for dimension reduction. stat 1050, 18 (2020).","journal-title":"stat"},{"key":"e_1_3_2_53_2","doi-asserted-by":"publisher","DOI":"10.1080\/00913367.2012.749090"},{"key":"e_1_3_2_54_2","doi-asserted-by":"publisher","DOI":"10.1007\/978-3-319-18461-6_52"},{"key":"e_1_3_2_55_2","doi-asserted-by":"publisher","DOI":"10.18653\/v1\/2020.acl-main.703"},{"key":"e_1_3_2_56_2","unstructured":"Alexey Dosovitskiy Lucas Beyer Alexander Kolesnikov Dirk Weissenborn Xiaohua Zhai Thomas Unterthiner Mostafa Dehghani Matthias Minderer Georg Heigold Sylvain Gelly Jakob Uszkoreit and Neil Houlsby. 2021. An image is worth 16x16 words: Transformers for image recognition at scale. In Proceedings of the International Conference on Learning Representations (ICLR\u201921)."},{"key":"e_1_3_2_57_2","unstructured":"Adam Paszke Sam Gross Francisco Massa Adam Lerer James Bradbury Gregory Chanan Trevor Killeen Zeming Lin Natalia Gimelshein Luca Antiga et\u00a0al. 2019. Pytorch: An imperative style high-performance deep learning library. In Proceedings of the 33rd International Conference on Neural Information Processing Systems. 8026\u20138037."},{"issue":"5","key":"e_1_3_2_58_2","first-page":"1","article-title":"Cpt: A pre-trained unbalanced transformer for both chinese language understanding and generation","volume":"67","author":"Shao Yunfan","year":"2024","unstructured":"Yunfan Shao, Zhichao Geng, Yitao Liu, Junqi Dai, Hang Yan, Fei Yang, Zhe Li, Hujun Bao, and Xipeng Qiu. 2024. Cpt: A pre-trained unbalanced transformer for both chinese language understanding and generation. Science China Information Sciences 67, 5 (2024), 1\u201313.","journal-title":"Science China Information Sciences"},{"key":"e_1_3_2_59_2","doi-asserted-by":"publisher","DOI":"10.18653\/v1\/N19-1423"},{"key":"e_1_3_2_60_2","unstructured":"Jinze Bai Shuai Bai Shusheng Yang Shijie Wang Sinan Tan Peng Wang Junyang Lin Chang Zhou and Jingren Zhou. 2023. Qwen-VL: A Frontier Large Vision-Language Model with Versatile Abilities. arXiv:2308.12966. Retrieved from https:\/\/arxiv.org\/abs\/2308.12966"},{"key":"e_1_3_2_61_2","unstructured":"Qinyuan Cheng Tianxiang Sun Wenwei Zhang Siyin Wang Xiangyang Liu Mozhi Zhang Junliang He Mianqiu Huang Zhangyue Yin Kai Chen and Xipeng Qiu. 2023. Evaluating Hallucinations in Chinese Large Language Models. arXiv:2310.03368. Retrieved from https:\/\/arxiv.org\/abs\/2310.03368"},{"key":"e_1_3_2_62_2","unstructured":"Haotian Liu Chunyuan Li Qingyang Wu and Yong Jae Lee. 2024. Visual instruction tuning. In Proceedings of the 37th International Conference on Neural Information Processing Systems. 34892\u201334916."},{"key":"e_1_3_2_63_2","doi-asserted-by":"publisher","DOI":"10.1007\/978-3-030-58452-8_13"},{"key":"e_1_3_2_64_2","volume-title":"Proceedings of the International Conference on Learning Representations","author":"Zhang Tianyi","year":"2020","unstructured":"Tianyi Zhang, Varsha Kishore, Felix Wu, Kilian Q. Weinberger, and Yoav Artzi. 2020. BERTScore: Evaluating text generation with BERT. In Proceedings of the International Conference on Learning Representations. Retrieved from https:\/\/openreview.net\/forum?id=SkeHuCVFDr"}],"container-title":["ACM Transactions on Asian and Low-Resource Language Information Processing"],"original-title":[],"language":"en","link":[{"URL":"https:\/\/dl.acm.org\/doi\/pdf\/10.1145\/3773989","content-type":"unspecified","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2025,12,10]],"date-time":"2025-12-10T14:24:51Z","timestamp":1765376691000},"score":1,"resource":{"primary":{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/3773989"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2025,12,10]]},"references-count":63,"journal-issue":{"issue":"12","published-print":{"date-parts":[[2025,12,31]]}},"alternative-id":["10.1145\/3773989"],"URL":"https:\/\/doi.org\/10.1145\/3773989","relation":{},"ISSN":["2375-4699","2375-4702"],"issn-type":[{"value":"2375-4699","type":"print"},{"value":"2375-4702","type":"electronic"}],"subject":[],"published":{"date-parts":[[2025,12,10]]},"assertion":[{"value":"2024-12-29","order":0,"name":"received","label":"Received","group":{"name":"publication_history","label":"Publication History"}},{"value":"2025-10-04","order":2,"name":"accepted","label":"Accepted","group":{"name":"publication_history","label":"Publication History"}},{"value":"2025-12-10","order":3,"name":"published","label":"Published","group":{"name":"publication_history","label":"Publication History"}}]}}