{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2026,8,3]],"date-time":"2026-08-03T07:03:17Z","timestamp":1785740597197,"version":"3.56.0"},"reference-count":23,"publisher":"Springer Science and Business Media LLC","issue":"8","license":[{"start":{"date-parts":[[2022,9,10]],"date-time":"2022-09-10T00:00:00Z","timestamp":1662768000000},"content-version":"tdm","delay-in-days":0,"URL":"https:\/\/www.springernature.com\/gp\/researchers\/text-and-data-mining"},{"start":{"date-parts":[[2022,9,10]],"date-time":"2022-09-10T00:00:00Z","timestamp":1662768000000},"content-version":"vor","delay-in-days":0,"URL":"https:\/\/www.springernature.com\/gp\/researchers\/text-and-data-mining"}],"funder":[{"DOI":"10.13039\/100007219","name":"Natural Science Foundation of Shanghai","doi-asserted-by":"publisher","award":["Grant No. 20ZR1420400"],"award-info":[{"award-number":["Grant No. 20ZR1420400"]}],"id":[{"id":"10.13039\/100007219","id-type":"DOI","asserted-by":"publisher"}]},{"DOI":"10.13039\/501100001809","name":"National Natural Science Foundation of China","doi-asserted-by":"publisher","award":["Grant No. 62172267"],"award-info":[{"award-number":["Grant No. 62172267"]}],"id":[{"id":"10.13039\/501100001809","id-type":"DOI","asserted-by":"publisher"}]},{"name":"Key Research Project of Zhejiang Laboratory","award":["No. 2021PE0AC02"],"award-info":[{"award-number":["No. 2021PE0AC02"]}]},{"DOI":"10.13039\/501100012166","name":"National Key R &D Program of China","doi-asserted-by":"crossref","award":["Grant No. 2019YFE0190500"],"award-info":[{"award-number":["Grant No. 2019YFE0190500"]}],"id":[{"id":"10.13039\/501100012166","id-type":"DOI","asserted-by":"crossref"}]},{"name":"State Key Program of National Natural Science Foundation of China","award":["Grant No. 61936001"],"award-info":[{"award-number":["Grant No. 61936001"]}]},{"name":"Shanghai Pujiang Program","award":["Grant No. 21PJ1404200"],"award-info":[{"award-number":["Grant No. 21PJ1404200"]}]}],"content-domain":{"domain":["link.springer.com"],"crossmark-restriction":false},"short-container-title":["J Ambient Intell Human Comput"],"published-print":{"date-parts":[[2023,8]]},"DOI":"10.1007\/s12652-022-04398-4","type":"journal-article","created":{"date-parts":[[2022,9,10]],"date-time":"2022-09-10T08:20:45Z","timestamp":1662798045000},"page":"11185-11194","update-policy":"https:\/\/doi.org\/10.1007\/springer_crossmark_policy","source":"Crossref","is-referenced-by-count":22,"title":["Multimodal contrastive learning for radiology report generation"],"prefix":"10.1007","volume":"14","author":[{"ORCID":"https:\/\/orcid.org\/0000-0001-5331-022X","authenticated-orcid":false,"given":"Xing","family":"Wu","sequence":"first","affiliation":[],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Jingwen","family":"Li","sequence":"additional","affiliation":[],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Jianjia","family":"Wang","sequence":"additional","affiliation":[],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Quan","family":"Qian","sequence":"additional","affiliation":[],"role":[{"vocabulary":"crossref","role":"author"}]}],"member":"297","published-online":{"date-parts":[[2022,9,10]]},"reference":[{"issue":"100","key":"4398_CR1","first-page":"557","volume":"24","author":"O Alfarghaly","year":"2021","unstructured":"Alfarghaly O, Khaled R, Elkorany A et al (2021) Automated radiology report generation using conditioned transformers. Inform Med Unlock 24(100):557","journal-title":"Inform Med Unlock"},{"key":"4398_CR2","doi-asserted-by":"crossref","unstructured":"Anderson P, He X, Buehler C, et\u00a0al (2018) Bottom\u2013up and top\u2013down attention for image captioning and visual question answering. In: 2018 IEEE\/CVF Conference on Computer Vision and Pattern Recognition, pp 6077\u20136086","DOI":"10.1109\/CVPR.2018.00636"},{"key":"4398_CR3","unstructured":"Banerjee S, Lavie A (2005) Meteor: An automatic metric for mt evaluation with improved correlation with human judgments. In: Proceedings of the Workshop on Intrinsic and Extrinsic Evaluation Measures for Machine Translation and\/or Summarization@ACL 2005, pp 65\u201372"},{"key":"4398_CR4","doi-asserted-by":"crossref","unstructured":"Chen Z, Song Y, Chang TH, et\u00a0al (2020) Generating radiology reports via memory-driven transformer. In: Proceedings of the 2020 Conference on Empirical Methods in Natural Language Processing, pp 1439\u20131449","DOI":"10.18653\/v1\/2020.emnlp-main.112"},{"issue":"2","key":"4398_CR5","doi-asserted-by":"publisher","first-page":"304","DOI":"10.1093\/jamia\/ocv080","volume":"23","author":"D Demner-Fushman","year":"2016","unstructured":"Demner-Fushman D, Kohli MD, Rosenman MB et al (2016) Preparing a collection of radiology examinations for distribution and retrieval. J Am Med Inform Assoc 23(2):304\u2013310","journal-title":"J Am Med Inform Assoc"},{"key":"4398_CR6","unstructured":"Devlin J, Chang M, Lee K, et\u00a0al (2019) BERT: pre-training of deep bidirectional transformers for language understanding. In: Proceedings of the 2019 Conference of the North American Chapter of the Association for Computational Linguistics: Human Language Technologies, pp 4171\u20134186"},{"key":"4398_CR7","doi-asserted-by":"crossref","unstructured":"Donahue J, Hendricks LA, Guadarrama S, et\u00a0al (2015) Long-term recurrent convolutional networks for visual recognition and description. In: 2015 IEEE Conference on Computer Vision and Pattern Recognition, pp 2625\u20132634","DOI":"10.1109\/CVPR.2015.7298878"},{"key":"4398_CR8","doi-asserted-by":"crossref","unstructured":"Gao T, Yao X, Chen D (2021) SimCSE: Simple contrastive learning of sentence embeddings. In: Proceedings of the 2021 Conference on Empirical Methods in Natural Language Processing, pp 6894\u20136910","DOI":"10.18653\/v1\/2021.emnlp-main.552"},{"key":"4398_CR9","doi-asserted-by":"crossref","unstructured":"Jing B, Xie P, Xing E (2018) On the automatic generation of medical imaging reports. In: Proceedings of the 56th Annual Meeting of the Association for Computational Linguistics, pp 2577\u20132586","DOI":"10.18653\/v1\/P18-1240"},{"key":"4398_CR10","doi-asserted-by":"publisher","unstructured":"Johnson AE, Pollard TJ, Greenbaum NR, et\u00a0al (2019) Mimic-cxr-jpg, a large publicly available database of labeled chest radiographs. Preprint at https:\/\/doi.org\/10.48550\/arXiv.1901.07042","DOI":"10.48550\/arXiv.1901.07042"},{"key":"4398_CR11","doi-asserted-by":"crossref","unstructured":"Krause J, Johnson J, Krishna R, et\u00a0al (2017) A hierarchical approach for generating descriptive image paragraphs. In: 2017 IEEE Conference on Computer Vision and Pattern Recognition (CVPR), pp 3337\u20133345","DOI":"10.1109\/CVPR.2017.356"},{"key":"4398_CR12","unstructured":"Li CY, Liang X, Hu Z, et\u00a0al (2018) Hybrid retrieval-generation reinforced agent for medical image report generation. In: Proceedings of the 32nd International Conference on Neural Information Processing Systems, p 1537\u20131547"},{"key":"4398_CR13","unstructured":"Lin CY (2004) Rouge: a package for automatic evaluation of summaries. In: Proceedings of the Workshop on Text Summarization Branches Out, Post-Conference Workshop of ACL 2004, Barcelona, Spain, pp 74\u201381"},{"key":"4398_CR14","doi-asserted-by":"crossref","unstructured":"Liu F, Ge S, Wu X (2021) Competence-based multimodal curriculum learning for medical report generation. In: Proceedings of the 59th Annual Meeting of the Association for Computational Linguistics and the 11th International Joint Conference on Natural Language Processing (Volume 1: Long Papers), pp 3001\u20133012","DOI":"10.18653\/v1\/2021.acl-long.234"},{"key":"4398_CR15","doi-asserted-by":"crossref","unstructured":"Lu J, Xiong C, Parikh D, et\u00a0al (2017) Knowing when to look: Adaptive attention via a visual sentinel for image captioning. In: 2017 IEEE Conference on Computer Vision and Pattern Recognition (CVPR), pp 3242\u20133250","DOI":"10.1109\/CVPR.2017.345"},{"key":"4398_CR16","doi-asserted-by":"crossref","unstructured":"Papineni K, Roukos S, Ward T, et\u00a0al (2002) Bleu: a method for automatic evaluation of machine translation. In: Proceedings of the 40th annual meeting of the Association for Computational Linguistics, p 311-318","DOI":"10.3115\/1073083.1073135"},{"key":"4398_CR17","doi-asserted-by":"crossref","unstructured":"Rennie SJ, Marcheret E, Mroueh Y, et\u00a0al (2017) Self-critical sequence training for image captioning. In: 2017 IEEE Conference on Computer Vision and Pattern Recognition (CVPR), pp 1179\u20131195","DOI":"10.1109\/CVPR.2017.131"},{"key":"4398_CR18","doi-asserted-by":"crossref","unstructured":"Vedantam R, Lawrence\u00a0Zitnick C, Parikh D (2015) Cider: consensus-based image description evaluation. In: 2015 IEEE Conference on Computer Vision and Pattern Recognition (CVPR), pp 4566\u20134575","DOI":"10.1109\/CVPR.2015.7299087"},{"key":"4398_CR19","doi-asserted-by":"crossref","unstructured":"Vinyals O, Toshev A, Bengio S, et\u00a0al (2015) Show and tell: a neural image caption generator. In: 2015 IEEE Conference on Computer Vision and Pattern Recognition (CVPR), pp 3156\u20133164","DOI":"10.1109\/CVPR.2015.7298935"},{"key":"4398_CR20","doi-asserted-by":"publisher","first-page":"457","DOI":"10.1007\/978-3-030-00928-1_52","volume":"2018","author":"Y Xue","year":"2018","unstructured":"Xue Y, Xu T, Rodney Long L et al (2018) Multimodal recurrent model with attention for automated radiology report generation. Med Image Comput Comput Assist Intervent MICCAI 2018:457\u2013466. https:\/\/doi.org\/10.1007\/978-3-030-00928-1_52","journal-title":"Med Image Comput Comput Assist Intervent MICCAI"},{"key":"4398_CR21","doi-asserted-by":"publisher","unstructured":"Yan Y, Li R, Wang S, et\u00a0al (2021) ConSERT: A contrastive framework for self-supervised sentence representation transfer. In: Proceedings of the 59th Annual Meeting of the Association for Computational Linguistics and the 11th International Joint Conference on Natural Language Processing, pp 5065\u20135075. https:\/\/doi.org\/10.18653\/v1\/2021.acl-long.393","DOI":"10.18653\/v1\/2021.acl-long.393"},{"key":"4398_CR22","doi-asserted-by":"publisher","unstructured":"Zhang Y, Jiang H, Miura Y, et\u00a0al (2020a) Contrastive learning of medical visual representations from paired images and text. Preprint at https:\/\/doi.org\/10.48550\/arXiv.2010.00747","DOI":"10.48550\/arXiv.2010.00747"},{"key":"4398_CR23","doi-asserted-by":"publisher","unstructured":"Zhang Y, Wang X, Xu Z, et\u00a0al (2020b) When radiology report generation meets knowledge graph. In: Proceedings of the AAAI Conference on Artificial Intelligence, pp 12910\u201312917. https:\/\/doi.org\/10.1609\/aaai.v34i07.6989","DOI":"10.1609\/aaai.v34i07.6989"}],"container-title":["Journal of Ambient Intelligence and Humanized Computing"],"original-title":[],"language":"en","link":[{"URL":"https:\/\/link.springer.com\/content\/pdf\/10.1007\/s12652-022-04398-4.pdf","content-type":"application\/pdf","content-version":"vor","intended-application":"text-mining"},{"URL":"https:\/\/link.springer.com\/article\/10.1007\/s12652-022-04398-4\/fulltext.html","content-type":"text\/html","content-version":"vor","intended-application":"text-mining"},{"URL":"https:\/\/link.springer.com\/content\/pdf\/10.1007\/s12652-022-04398-4.pdf","content-type":"application\/pdf","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2023,6,21]],"date-time":"2023-06-21T14:17:45Z","timestamp":1687357065000},"score":1,"resource":{"primary":{"URL":"https:\/\/link.springer.com\/10.1007\/s12652-022-04398-4"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2022,9,10]]},"references-count":23,"journal-issue":{"issue":"8","published-print":{"date-parts":[[2023,8]]}},"alternative-id":["4398"],"URL":"https:\/\/doi.org\/10.1007\/s12652-022-04398-4","relation":{},"ISSN":["1868-5137","1868-5145"],"issn-type":[{"value":"1868-5137","type":"print"},{"value":"1868-5145","type":"electronic"}],"subject":[],"published":{"date-parts":[[2022,9,10]]},"assertion":[{"value":"25 April 2022","order":1,"name":"received","label":"Received","group":{"name":"ArticleHistory","label":"Article History"}},{"value":"1 September 2022","order":2,"name":"accepted","label":"Accepted","group":{"name":"ArticleHistory","label":"Article History"}},{"value":"10 September 2022","order":3,"name":"first_online","label":"First Online","group":{"name":"ArticleHistory","label":"Article History"}}]}}