{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2026,4,25]],"date-time":"2026-04-25T11:42:49Z","timestamp":1777117369425,"version":"3.51.4"},"reference-count":38,"publisher":"Springer Science and Business Media LLC","issue":"6","license":[{"start":{"date-parts":[[2026,4,25]],"date-time":"2026-04-25T00:00:00Z","timestamp":1777075200000},"content-version":"tdm","delay-in-days":0,"URL":"https:\/\/www.springernature.com\/gp\/researchers\/text-and-data-mining"},{"start":{"date-parts":[[2026,4,25]],"date-time":"2026-04-25T00:00:00Z","timestamp":1777075200000},"content-version":"vor","delay-in-days":0,"URL":"https:\/\/www.springernature.com\/gp\/researchers\/text-and-data-mining"}],"funder":[{"DOI":"10.13039\/501100007620","name":"Department of Education of Liaoning Province","doi-asserted-by":"crossref","award":["JYTMS20230012"],"award-info":[{"award-number":["JYTMS20230012"]}],"id":[{"id":"10.13039\/501100007620","id-type":"DOI","asserted-by":"crossref"}]}],"content-domain":{"domain":["link.springer.com"],"crossmark-restriction":false},"short-container-title":["J Supercomput"],"DOI":"10.1007\/s11227-026-08547-w","type":"journal-article","created":{"date-parts":[[2026,4,25]],"date-time":"2026-04-25T11:17:04Z","timestamp":1777115824000},"update-policy":"https:\/\/doi.org\/10.1007\/springer_crossmark_policy","source":"Crossref","is-referenced-by-count":0,"title":["KGRA: A knowledge-guided and relation-aware model for enhanced multimodal relation extraction"],"prefix":"10.1007","volume":"82","author":[{"given":"Wei","family":"Zheng","sequence":"first","affiliation":[],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Guoyin","family":"Li","sequence":"additional","affiliation":[],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Zhenlin","family":"Zhang","sequence":"additional","affiliation":[],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Hongfei","family":"Lin","sequence":"additional","affiliation":[],"role":[{"role":"author","vocabulary":"crossref"}]}],"member":"297","published-online":{"date-parts":[[2026,4,25]]},"reference":[{"key":"8547_CR1","unstructured":"Chen F, Feng Y (2023) Chain-of-thought prompt distillation for multimodal named entity recognition and multimodal relation extraction. arXiv preprint arXiv:2306.14122."},{"key":"8547_CR2","doi-asserted-by":"crossref","unstructured":"Chen X, Zhang N, Li L, Yao Y, Deng S, Tan C, ... Chen H (2022) Good visual guidance makes a better extractor: Hierarchical visual prefix for multimodal entity and relation extraction. arXiv 2022. arXiv preprint arXiv:2205.03521.","DOI":"10.18653\/v1\/2022.findings-naacl.121"},{"key":"8547_CR3","doi-asserted-by":"crossref","unstructured":"Chen X, Zhang N, Li L, Deng S, Tan C, Xu C, ... Chen H (2022). Hybrid transformer with multi-level fusion for multimodal knowledge graph completion. IN PROCEEDINGS OF THE 45TH INTERNATIONAL ACM SIGIR CONFERENCE ON RESEARCH AND DEVELOPMENT IN INFORMATION RETRIEVAL (pp. 904\u2013915).","DOI":"10.1145\/3477495.3531992"},{"key":"8547_CR4","doi-asserted-by":"publisher","first-page":"1274","DOI":"10.1109\/TASLP.2023.3345146","volume":"32","author":"S Cui","year":"2024","unstructured":"Cui S, Cao J, Cong X et al (2024) Enhancing multimodal entity and relation extraction with variational information bottleneck. IEEE\/ACM Trans Audio, Speech, Lang Process 32:1274\u20131285","journal-title":"IEEE\/ACM Trans Audio, Speech, Lang Process"},{"key":"8547_CR5","doi-asserted-by":"crossref","unstructured":"Devlin J, Chang MW, Lee K, Toutanova K (2019). Bert: pre-training of deep bidirectional transformers for language understanding. IN PROCEEDINGS OF THE 2019 CONFERENCE OF THE NORTH AMERICAN CHAPTER OF THE ASSOCIATION FOR COMPUTATIONAL LINGUISTICS: HUMAN LANGUAGE TECHNOLOGIES, vol 1 (long and short papers) (pp. 4171\u20134186).","DOI":"10.18653\/v1\/N19-1423"},{"key":"8547_CR6","unstructured":"Dosovitskiy A, Beyer L, Kolesnikov A, Weissenborn D, Zhai X, Unterthiner T Houlsby N (2020) An image is worth 16x16 words: transformers for image recognition at scale. arXiv preprint arXiv:2010.11929."},{"key":"8547_CR7","doi-asserted-by":"publisher","first-page":"125608","DOI":"10.1016\/j.eswa.2024.125608","volume":"262","author":"Y Gong","year":"2025","unstructured":"Gong Y, Lv X, Yuan Z et al (2025) CE-DCVSI: multimodal relational extraction based on collaborative enhancement of dual-channel visual semantic information. Expert Syst Appl 262:125608","journal-title":"Expert Syst Appl"},{"key":"8547_CR8","doi-asserted-by":"crossref","unstructured":"Gong P, Liu J, Zhang X, Li X (2023) A multi-stage hierarchical relational graph neural network for multimodal sentiment analysis. In ICASSP 2023\u20132023 IEEE INTERNATIONAL CONFERENCE ON ACOUSTICS, SPEECH AND SIGNAL PROCESSING (ICASSP) (pp. 1\u20135). IEEE.","DOI":"10.1109\/ICASSP49357.2023.10096644"},{"issue":"22","key":"8547_CR9","doi-asserted-by":"publisher","first-page":"12208","DOI":"10.3390\/app132212208","volume":"13","author":"W He","year":"2023","unstructured":"He W, Ma H, Li S et al (2023) Using augmented small multimodal models to guide large language models for multimodal relation extraction. Appl Sci 13(22):12208","journal-title":"Appl Sci"},{"issue":"1","key":"8547_CR10","doi-asserted-by":"publisher","DOI":"10.1016\/j.ipm.2024.103875","volume":"62","author":"X He","year":"2025","unstructured":"He X, Li S, Zhang Y, Li B, Xu S, Zhou Y (2025) The more quality information the better: Hierarchical generation of multi-evidence alignment and fusion model for multimodal entity and relation extraction. Inf Process Manage 62(1):103875","journal-title":"Inf Process Manage"},{"key":"8547_CR11","doi-asserted-by":"crossref","unstructured":"Hu X, Chen J, Liu A, Meng S, Wen L, Yu PS (2023) Prompt me up: unleashing the power of alignments for multimodal entity and relation extraction. IN PROCEEDINGS OF THE 31ST ACM INTERNATIONAL CONFERENCE ON MULTIMEDIA (pp. 5185\u20135194)","DOI":"10.1145\/3581783.3611899"},{"key":"8547_CR12","doi-asserted-by":"crossref","unstructured":"Hu X, Guo Z, Teng Z, King I, Yu PS (2023) Multimodal relation extraction with cross-modal retrieval and synthesis. arXiv preprint arXiv:2305.16166.","DOI":"10.18653\/v1\/2023.acl-short.27"},{"key":"8547_CR13","doi-asserted-by":"crossref","unstructured":"Hu Z, Dong Y, Wang K, Sun Y (2020) Heterogeneous graph transformer. IN PROCEEDINGS OF THE WEB CONFERENCE 2020 (pp. 2704\u20132710)","DOI":"10.1145\/3366423.3380027"},{"issue":"3","key":"8547_CR14","doi-asserted-by":"publisher","DOI":"10.1016\/j.ipm.2024.104033","volume":"62","author":"S Huang","year":"2025","unstructured":"Huang S, Cai Y, Yuan L, Wang J (2025) A knowledge-enhanced network for joint multimodal entity-relation extraction. Inf Process Manage 62(3):104033","journal-title":"Inf Process Manage"},{"key":"8547_CR15","doi-asserted-by":"crossref","unstructured":"Huang Y, Lin Z (2023). I2SRM: Intra-and Inter-Sample Relationship Modeling for Multimodal Information Extraction. IN PROCEEDINGS OF THE 5TH ACM INTERNATIONAL CONFERENCE ON MULTIMEDIA IN ASIA (pp. 1\u20131)","DOI":"10.1145\/3595916.3626399"},{"key":"8547_CR16","doi-asserted-by":"crossref","unstructured":"Jia Z, Lin Y, Wang J, Feng Z, Xie X, Chen C (2021) HetEmotionNet: two-stream heterogeneous graph recurrent neural network for multi-modal emotion recognition. IN PROCEEDINGS OF THE 29TH ACM INTERNATIONAL CONFERENCE ON MULTIMEDIA (pp. 1047\u20131056)","DOI":"10.1145\/3474085.3475583"},{"key":"8547_CR17","unstructured":"Kipf TN, Welling M (2016) Semi-supervised classification with graph convolutional networks. arXiv preprint arXiv:1609.02907"},{"key":"8547_CR18","doi-asserted-by":"crossref","unstructured":"Li Q, Guo S, Ji C, Peng X, Cui S, Li J (2023) Dual-gated fusion with prefix-tuning for multi-modal relation extraction. arXiv preprint arXiv:2306.11020","DOI":"10.18653\/v1\/2023.findings-acl.572"},{"key":"8547_CR19","unstructured":"Li J, Li D, Xiong C, Hoi S (2022) Blip: bootstrapping language-image pre-training for unified vision-language understanding and generation. IN INTERNATIONAL CONFERENCE ON MACHINE LEARNING (pp. 12888\u201312900). PMLR"},{"key":"8547_CR20","doi-asserted-by":"publisher","first-page":"119815","DOI":"10.1016\/j.ins.2023.119815","volume":"654","author":"J Li","year":"2024","unstructured":"Li J, Yang C, Ye G et al (2024) Graph neural networks with deep mutual learning for designing multi-modal recommendation systems. Info Sci 654:119815","journal-title":"Info Sci"},{"key":"8547_CR21","doi-asserted-by":"crossref","unstructured":"Li L, Chen,X, Qiao S, et al (2023) On analyzing the role of image for visual enhanced relation extraction. IN PROCEEDINGS OF THE 37TH AAAI CONFERENCE ON ARTIFICIAL INTELLIGENCE AND 35TH CONFERENCE ON INNOVATIVE APPLICATIONS OF ARTIFICIAL INTELLIGENCE AND 13TH SYMPOSIUM ON EDUCATIONAL ADVANCES IN ARTIFICIAL INTELLIGENCE (pp. 16254\u201316255).","DOI":"10.1609\/aaai.v37i13.26987"},{"issue":"2","key":"8547_CR22","doi-asserted-by":"publisher","first-page":"578","DOI":"10.1093\/bib\/bbab578","volume":"23","author":"S Mahbub","year":"2022","unstructured":"Mahbub S, Bayzid MS (2022) EGRET: edge aggregated graph attention networks and transfer learning improve protein\u2013protein interaction site prediction. Brief Bioinform 23(2):578","journal-title":"Brief Bioinform"},{"key":"8547_CR23","unstructured":"Soares LB, FitzGerald N, Ling J, Kwiatkowski T (2019) Matching the blanks: distributional similarity for relation learning. arXiv preprint arXiv:1906.03158"},{"issue":"20","key":"8547_CR24","first-page":"10","volume":"1050","author":"P Velickovic","year":"2017","unstructured":"Velickovic P, Cucurull G, Casanova A, Romero A, Lio P, Bengio Y (2017) Graph atten networks stat 1050(20):10\u201348550","journal-title":"Graph atten networks stat"},{"key":"8547_CR25","doi-asserted-by":"publisher","DOI":"10.7717\/peerj-cs.1856","volume":"10","author":"M Wang","year":"2024","unstructured":"Wang M, Chen H, Shen D, Li B, Hu S (2024) RSRNeT: a novel multi-modal network framework for named entity recognition and relation extraction. PeerJ Comput Sci 10:e1856","journal-title":"PeerJ Comput Sci"},{"key":"8547_CR26","doi-asserted-by":"crossref","unstructured":"Wang X, Cai J, Jiang Y, Xie P, Tu K, Lu W (2022) Named entity and relation extraction with multi-modal retrieval. arXiv preprint arXiv:2212.01612","DOI":"10.18653\/v1\/2022.findings-emnlp.437"},{"key":"8547_CR27","doi-asserted-by":"crossref","unstructured":"Wang Y, Yasunaga M, Ren H, Wada S, Leskovec J (2023) Vqa-gnn: reasoning with multimodal knowledge via graph neural networks for visual question answering. IN PROCEEDINGS OF THE IEEE\/CVF INTERNATIONAL CONFERENCE ON COMPUTER VISION (pp. 21582\u201321592).","DOI":"10.1109\/ICCV51070.2023.01973"},{"key":"8547_CR28","doi-asserted-by":"crossref","unstructured":"Wu S, Fei H, Cao Y, Bing L, Chua TS (2023) Information screening whilst exploiting! multimodal relation extraction with feature denoising and multimodal topic modeling. arXiv preprint arXiv:2305.11719","DOI":"10.18653\/v1\/2023.acl-long.823"},{"issue":"3","key":"8547_CR29","doi-asserted-by":"publisher","first-page":"1","DOI":"10.1145\/3450352","volume":"39","author":"T Yang","year":"2021","unstructured":"Yang T, Hu L, Shi C et al (2021) HGAT: Heterogeneous graph attention networks for semi-supervised short text classification. ACM Trans Inform Syst 39(3):1\u201329","journal-title":"ACM Trans Inform Syst"},{"key":"8547_CR30","doi-asserted-by":"crossref","unstructured":"Yang Z, Gong B, Wang L, Huang W, Yu D, Luo J (2019) A fast and accurate one-stage approach to visual grounding. IN PROCEEDINGS OF THE IEEE\/CVF INTERNATIONAL CONFERENCE ON COMPUTER VISION (pp. 4683\u20134693)","DOI":"10.1109\/ICCV.2019.00478"},{"key":"8547_CR31","doi-asserted-by":"publisher","first-page":"103986","DOI":"10.1016\/j.artint.2023.103986","volume":"32","author":"Y Yin","year":"2023","unstructured":"Yin Y, Zeng J, Su J et al (2023) Multi-modal graph contrastive encoding for neural machine translation. Artif Intell 32:103986","journal-title":"Artif Intell"},{"key":"8547_CR32","doi-asserted-by":"crossref","unstructured":"Yuan L, Cai Y, Wang J, et al (2023) Joint multimodal entity relation extraction based on edge-enhanced graph alignment network and word-pair relation tagging. IN PROCEEDINGS OF THE 37TH AAAI CONFERENCE ON ARTIFICIAL INTELLIGENCE AND 35TH CONFERENCE ON INNOVATIVE APPLICATIONS OF ARTIFICIAL INTELLIGENCE AND 13TH SYMPOSIUM ON EDUCATIONAL ADVANCES IN ARTIFICIAL INTELLIGENCE (pp. 11051\u201311059).","DOI":"10.1609\/aaai.v37i9.26309"},{"key":"8547_CR33","doi-asserted-by":"crossref","unstructured":"Zeng D, Liu K, Chen Y, Zhao J (2015) Distant supervision for relation extraction via piecewise convolutional neural networks. IN PROCEEDINGS OF THE 2015 CONFERENCE ON EMPIRICAL METHODS IN NATURAL LANGUAGE PROCESSING (pp. 1753\u20131762)","DOI":"10.18653\/v1\/D15-1203"},{"issue":"3","key":"8547_CR34","doi-asserted-by":"publisher","first-page":"103264","DOI":"10.1016\/j.ipm.2023.103264","volume":"60","author":"Q Zhao","year":"2023","unstructured":"Zhao Q, Gao T, Guo N (2023) TSVFN: two-stage visual fusion network for multimodal relation extraction. Inform Process Manag 60(3):103264","journal-title":"Inform Process Manag"},{"key":"8547_CR35","doi-asserted-by":"crossref","unstructured":"Zheng C, Feng J, Fu Z, Cai Y, Li Q, Wang T (2021) Multimodal relation extraction with efficient graph alignment. IN PROCEEDINGS OF THE 29TH ACM INTERNATIONAL CONFERENCE ON MULTIMEDIA (pp. 5298\u20135306)","DOI":"10.1145\/3474085.3476968"},{"key":"8547_CR36","doi-asserted-by":"crossref","unstructured":"Zheng C, Wu Z, Feng J, Fu Z, Cai Y (2021) MNRE: a challenge multimodal dataset for neural relation extraction with visual evidence in social media posts. IN 2021 IEEE INTERNATIONAL CONFERENCE ON MULTIMEDIA AND EXPO (ICME) (pp. 1\u20136). IEEE","DOI":"10.1109\/ICME51207.2021.9428274"},{"key":"8547_CR37","doi-asserted-by":"crossref","unstructured":"Zhong Z, Chen D (2020) A frustratingly easy approach for entity and relation extraction. arXiv preprint arXiv:2010.12812","DOI":"10.18653\/v1\/2021.naacl-main.5"},{"issue":"10","key":"8547_CR38","doi-asserted-by":"publisher","first-page":"6178","DOI":"10.3390\/app13106178","volume":"13","author":"M Zuo","year":"2023","unstructured":"Zuo M, Wang Y, Dong W et al (2023) Visual description augmented integration network for multimodal entity and relation extraction. Appl Sci 13(10):6178","journal-title":"Appl Sci"}],"container-title":["The Journal of Supercomputing"],"original-title":[],"language":"en","link":[{"URL":"https:\/\/link.springer.com\/content\/pdf\/10.1007\/s11227-026-08547-w.pdf","content-type":"application\/pdf","content-version":"vor","intended-application":"text-mining"},{"URL":"https:\/\/link.springer.com\/article\/10.1007\/s11227-026-08547-w","content-type":"text\/html","content-version":"vor","intended-application":"text-mining"},{"URL":"https:\/\/link.springer.com\/content\/pdf\/10.1007\/s11227-026-08547-w.pdf","content-type":"application\/pdf","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2026,4,25]],"date-time":"2026-04-25T11:17:12Z","timestamp":1777115832000},"score":1,"resource":{"primary":{"URL":"https:\/\/link.springer.com\/10.1007\/s11227-026-08547-w"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2026,4,25]]},"references-count":38,"journal-issue":{"issue":"6","published-online":{"date-parts":[[2026,4]]}},"alternative-id":["8547"],"URL":"https:\/\/doi.org\/10.1007\/s11227-026-08547-w","relation":{},"ISSN":["1573-0484"],"issn-type":[{"value":"1573-0484","type":"electronic"}],"subject":[],"published":{"date-parts":[[2026,4,25]]},"assertion":[{"value":"7 November 2025","order":1,"name":"received","label":"Received","group":{"name":"ArticleHistory","label":"Article History"}},{"value":"18 April 2026","order":2,"name":"accepted","label":"Accepted","group":{"name":"ArticleHistory","label":"Article History"}},{"value":"25 April 2026","order":3,"name":"first_online","label":"First Online","group":{"name":"ArticleHistory","label":"Article History"}},{"order":1,"name":"Ethics","group":{"name":"EthicsHeading","label":"Declarations"}},{"value":"The authors declare that they have no known conflict of financial interest or personal relationships that could have appeared to influence the work reported in this paper.","order":2,"name":"Ethics","group":{"name":"EthicsHeading","label":"Conflict of interest"}}],"article-number":"378"}}