{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2025,11,13]],"date-time":"2025-11-13T07:22:42Z","timestamp":1763018562444,"version":"3.41.0"},"reference-count":50,"publisher":"Association for Computing Machinery (ACM)","issue":"6","license":[{"start":{"date-parts":[[2023,5,30]],"date-time":"2023-05-30T00:00:00Z","timestamp":1685404800000},"content-version":"vor","delay-in-days":0,"URL":"https:\/\/www.acm.org\/publications\/policies\/copyright_policy#Background"}],"content-domain":{"domain":["dl.acm.org"],"crossmark-restriction":true},"short-container-title":["ACM Trans. Multimedia Comput. Commun. Appl."],"published-print":{"date-parts":[[2023,11,30]]},"abstract":"<jats:p>Visual paraphrase generation task aims to rewrite a given image-related original sentence into a new paraphrase, where the paraphrase needs to have the same expressed meaning as the original sentence but have a difference in expression form. Existing studies mainly extract two semantic vectors to represent the entire image and the entire original sentence, respectively, for paraphrase generation. However, these semantic vectors for an image or a sentence may lead to the model failing to focus on some key objects in the original sentence, which may generate semantically inconsistent sentences by changing key object information. In this article, we propose an object-level paraphrase generation model, which generates paraphrases by adjusting the permutation of key objects and modifying their associated descriptions. To adjust the permutation of key objects, an object-sorting module aims to obtain new object sequences based on the key object information and original sentences. Then, a sequence generation module sequentially generates paraphrases based on the permutation of the newly object sequences. Each generation step focuses on different image features associated with different key objects to generate descriptions with differences. Furthermore, we use a semantic discriminator module to promote the generated paraphrase to be semantically close to the original sentence. Specifically, the loss function of the discriminator penalizes the excessive distance between the paraphrase and the original sentence. Extensive experiments on the MS COCO dataset show that the proposed model outperforms the baselines.<\/jats:p>\n          <jats:p\/>","DOI":"10.1145\/3585010","type":"journal-article","created":{"date-parts":[[2023,2,27]],"date-time":"2023-02-27T12:04:13Z","timestamp":1677499453000},"page":"1-19","update-policy":"https:\/\/doi.org\/10.1145\/crossmark-policy","source":"Crossref","is-referenced-by-count":2,"title":["Visual Paraphrase Generation with Key Information Retained"],"prefix":"10.1145","volume":"19","author":[{"ORCID":"https:\/\/orcid.org\/0000-0002-6833-7879","authenticated-orcid":false,"given":"Jiayuan","family":"Xie","sequence":"first","affiliation":[{"name":"School of Software Engineering, South China University of Technology, Guangzhou, China and Key Laboratory of Big Data and Intelligent Robot (SCUT), MOE of China"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"ORCID":"https:\/\/orcid.org\/0000-0001-8064-1577","authenticated-orcid":false,"given":"Jiali","family":"Chen","sequence":"additional","affiliation":[{"name":"School of Software Engineering, South China University of Technology, Guangzhou, China and Key Laboratory of Big Data and Intelligent Robot (SCUT), MOE of China"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"ORCID":"https:\/\/orcid.org\/0000-0002-1767-789X","authenticated-orcid":false,"given":"Yi","family":"Cai","sequence":"additional","affiliation":[{"name":"School of Software Engineering, South China University of Technology, Guangzhou, China and Key Laboratory of Big Data and Intelligent Robot (SCUT), MOE of China and Peng Cheng Laboratory, Shenzhen, China"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"ORCID":"https:\/\/orcid.org\/0000-0001-7691-347X","authenticated-orcid":false,"given":"Qingbao","family":"Huang","sequence":"additional","affiliation":[{"name":"School of Electrical Engineering, Guangxi University, Nanning, Guangxi, China"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"ORCID":"https:\/\/orcid.org\/0000-0003-3370-471X","authenticated-orcid":false,"given":"Qing","family":"Li","sequence":"additional","affiliation":[{"name":"Department of Computing, The Hong Kong Polytechnic University, Hong Kong"}],"role":[{"role":"author","vocabulary":"crossref"}]}],"member":"320","published-online":{"date-parts":[[2023,5,30]]},"reference":[{"key":"e_1_3_1_2_2","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR.2018.00636"},{"key":"e_1_3_1_3_2","doi-asserted-by":"crossref","first-page":"300","DOI":"10.18653\/v1\/2022.acl-srw.23","volume-title":"Proceedings of the 60th Annual Meeting of the Association for Computational Linguistics: Student Research Workshop (ACL\u201922)","author":"Babakov Nikolay","year":"2022","unstructured":"Nikolay Babakov, David Dale, Varvara Logacheva, and Alexander Panchenko. 2022. A large-scale computational study of content preservation measures for text style transfer and paraphrase generation. In Proceedings of the 60th Annual Meeting of the Association for Computational Linguistics: Student Research Workshop (ACL\u201922), Samuel Louvan, Andrea Madotto, and Brielen Madureira (Eds.). Association for Computational Linguistics, 300\u2013321."},{"key":"e_1_3_1_4_2","first-page":"596","volume-title":"Proceedings of the 60th Annual Meeting of the Association for Computational Linguistics (ACL\u201922)","author":"Bandel Elron","year":"2022","unstructured":"Elron Bandel, Ranit Aharonov, Michal Shmueli-Scheuer, Ilya Shnayderman, Noam Slonim, and Liat Ein-Dor. 2022. Quality controlled paraphrase generation. In Proceedings of the 60th Annual Meeting of the Association for Computational Linguistics (ACL\u201922), Smaranda Muresan, Preslav Nakov, and Aline Villavicencio (Eds.). Association for Computational Linguistics, 596\u2013609."},{"key":"e_1_3_1_5_2","volume-title":"Proceedings of the 8th International Conference on Learning Representations (ICLR\u201920)","author":"Cai Han","year":"2020","unstructured":"Han Cai, Chuang Gan, Tianzhe Wang, Zhekai Zhang, and Song Han. 2020. Once-for-all: Train one network and specialize it for efficient deployment. In Proceedings of the 8th International Conference on Learning Representations (ICLR\u201920)."},{"key":"e_1_3_1_6_2","doi-asserted-by":"crossref","first-page":"376","DOI":"10.3115\/v1\/W14-3348","volume-title":"Proceedings of the 9th Workshop on Statistical Machine Translation (ACL\u201914)","author":"Denkowski Michael J.","year":"2014","unstructured":"Michael J. Denkowski and Alon Lavie. 2014. Meteor universal: Language specific translation evaluation for any target language. In Proceedings of the 9th Workshop on Statistical Machine Translation (ACL\u201914). 376\u2013380."},{"key":"e_1_3_1_7_2","unstructured":"Jacob Devlin Ming-Wei Chang Kenton Lee and Kristina Toutanova. 2018. Bert: Pre-training of deep bidirectional transformers for language understanding. Retrieved from https:\/\/arXiv:1810.04805."},{"key":"e_1_3_1_8_2","first-page":"2625","volume-title":"Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition (CVPR\u201915)","author":"Donahue Jeff","year":"2015","unstructured":"Jeff Donahue, Lisa Anne Hendricks, Sergio Guadarrama, Marcus Rohrbach, Subhashini Venugopalan, Trevor Darrell, and Kate Saenko. 2015. Long-term recurrent convolutional networks for visual recognition and description. In Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition (CVPR\u201915). 2625\u20132634."},{"key":"e_1_3_1_9_2","doi-asserted-by":"publisher","DOI":"10.18653\/v1\/D18-1045"},{"key":"e_1_3_1_10_2","doi-asserted-by":"crossref","unstructured":"Marvin Eisenberger Aysim Toker Laura Leal-Taix\u00e9 Florian Bernard and Daniel Cremers. 2022. A unified framework for implicit sinkhorn differentiation. In Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition (CVPR\u201922) .IEEE 499\u2013508.","DOI":"10.1109\/CVPR52688.2022.00059"},{"key":"e_1_3_1_11_2","first-page":"657","volume-title":"Proceedings of the International Joint Conference on Artificial Intelligence (IJCAI\u201921)","author":"Fan Zhihao","year":"2021","unstructured":"Zhihao Fan, Zhongyu Wei, Siyuan Wang, Ruize Wang, Zejun Li, Haijun Shan, and Xuanjing Huang. 2021. TCIC: Theme concepts learning cross language and vision for image captioning. In Proceedings of the International Joint Conference on Artificial Intelligence (IJCAI\u201921), Zhi-Hua Zhou (Ed.). ijcai.org, 657\u2013663."},{"key":"e_1_3_1_12_2","first-page":"1473","volume-title":"Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition (CVPR\u201915)","author":"Fang Hao","year":"2015","unstructured":"Hao Fang, Saurabh Gupta, Forrest N. Iandola, Rupesh Kumar Srivastava, Li Deng, Piotr Doll\u00e1r, Jianfeng Gao, Xiaodong He, Margaret Mitchell, John C. Platt, C. Lawrence Zitnick, and Geoffrey Zweig. 2015. From captions to visual concepts and back. In Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition (CVPR\u201915). 1473\u20131482."},{"issue":"2","key":"e_1_3_1_13_2","first-page":"50:1\u201350:21","article-title":"BTDP: Toward sparse fusion with block term decomposition pooling for visual question answering","volume":"15","author":"Fang Zhiwei","year":"2019","unstructured":"Zhiwei Fang, Jing Liu, Xueliang Liu, Qu Tang, Yong Li, and Hanqing Lu. 2019. BTDP: Toward sparse fusion with block term decomposition pooling for visual question answering. ACM Trans. Multim. Comput. Commun. Appl. 15, 2s (2019), 50:1\u201350:21.","journal-title":"ACM Trans. Multim. Comput. Commun. Appl."},{"key":"e_1_3_1_14_2","first-page":"238","volume-title":"Proceedings of the Annual Meeting of the Association for Computational Linguistics (ACL\u201920)","author":"Goyal Tanya","year":"2020","unstructured":"Tanya Goyal and Greg Durrett. 2020. Neural syntactic preordering for controlled paraphrase generation. In Proceedings of the Annual Meeting of the Association for Computational Linguistics (ACL\u201920). 238\u2013252."},{"key":"e_1_3_1_15_2","doi-asserted-by":"publisher","DOI":"10.1162\/neco.1997.9.8.1735"},{"key":"e_1_3_1_16_2","first-page":"2489","volume-title":"Proceedings of the 60th Annual Meeting of the Association for Computational Linguistics (ACL\u201922)","author":"Hosking Tom","year":"2022","unstructured":"Tom Hosking, Hao Tang, and Mirella Lapata. 2022. Hierarchical sketch induction for paraphrase generation. In Proceedings of the 60th Annual Meeting of the Association for Computational Linguistics (ACL\u201922). Association for Computational Linguistics, 2489\u20132501."},{"key":"e_1_3_1_17_2","first-page":"1419","volume-title":"Proceedings of the IEEE\/CVF International Conference on Computer Vision (ICCV\u201921)","author":"Hu Ronghang","year":"2021","unstructured":"Ronghang Hu and Amanpreet Singh. 2021. UniT: Multimodal multitask learning with a unified transformer. In Proceedings of the IEEE\/CVF International Conference on Computer Vision (ICCV\u201921). 1419\u20131429."},{"key":"e_1_3_1_18_2","unstructured":"Zhiheng Huang Wei Xu and Kai Yu. 2015. Bidirectional LSTM-CRF models for sequence tagging. Retrieved from https:\/\/abs\/1508.01991."},{"issue":"2","key":"e_1_3_1_19_2","first-page":"52:1\u201352:22","article-title":"Video question answering via knowledge-based progressive spatial-temporal attention network","volume":"15","author":"Jin Weike","year":"2019","unstructured":"Weike Jin, Zhou Zhao, Yimeng Li, Jie Li, Jun Xiao, and Yueting Zhuang. 2019. Video question answering via knowledge-based progressive spatial-temporal attention network. ACM Trans. Multim. Comput. Commun. Appl. 15, 2s (2019), 52:1\u201352:22.","journal-title":"ACM Trans. Multim. Comput. Commun. Appl."},{"key":"e_1_3_1_20_2","first-page":"1591","volume-title":"Proceeding of the INTERSPEECH","author":"Kim Jaeyoung","year":"2017","unstructured":"Jaeyoung Kim, Mostafa El-Khamy, and Jungwon Lee. 2017. Residual LSTM: Design of a deep recurrent architecture for distant speech recognition. In Proceeding of the INTERSPEECH, Francisco Lacerda (Ed.). ISCA, 1591\u20131595."},{"key":"e_1_3_1_21_2","volume-title":"Proceedings of the 3rd International Conference on Learning Representations (ICLR\u201915)","author":"Kingma Diederik P.","year":"2015","unstructured":"Diederik P. Kingma and Jimmy Ba. 2015. Adam: A method for stochastic optimization. In Proceedings of the 3rd International Conference on Learning Representations (ICLR\u201915)."},{"key":"e_1_3_1_22_2","first-page":"595","volume-title":"Proceedings of the 31st International Conference on Machine Learning (ICML\u201914)","author":"Kiros Ryan","year":"2014","unstructured":"Ryan Kiros, Ruslan Salakhutdinov, and Rich Zemel. 2014. Multimodal neural language models. In Proceedings of the 31st International Conference on Machine Learning (ICML\u201914). 595\u2013603."},{"key":"e_1_3_1_23_2","doi-asserted-by":"publisher","DOI":"10.1007\/s11263-016-0981-7"},{"key":"e_1_3_1_24_2","doi-asserted-by":"crossref","unstructured":"Liunian Harold Li Haoxuan You Zhecan Wang Alireza Zareian Shih-Fu Chang and Kai-Wei Chang. 2020. Weakly-supervised VisualBERT: Pre-training without parallel images and captions. Retrieved from https:\/\/abs\/2010.12831.","DOI":"10.18653\/v1\/2021.naacl-main.420"},{"issue":"3","key":"e_1_3_1_25_2","first-page":"76:1\u201376:21","article-title":"Inner knowledge-based Img2Doc scheme for visual question answering","volume":"18","author":"Li Qun","year":"2022","unstructured":"Qun Li, Fu Xiao, Bir Bhanu, Biyun Sheng, and Richang Hong. 2022. Inner knowledge-based Img2Doc scheme for visual question answering. ACM Trans. Multim. Comput. Commun. Appl. 18, 3 (2022), 76:1\u201376:21.","journal-title":"ACM Trans. Multim. Comput. Commun. Appl."},{"key":"e_1_3_1_26_2","first-page":"2592","volume-title":"Proceedings of the 59th Annual Meeting of the Association for Computational Linguistics and the 11th International Joint Conference on Natural Language Processing (ACL\/IJCNLP\u201921)","author":"Li Wei","year":"2021","unstructured":"Wei Li, Can Gao, Guocheng Niu, Xinyan Xiao, Hao Liu, Jiachen Liu, Hua Wu, and Haifeng Wang. 2021. UNIMO: Towards unified-modal understanding and generation via cross-modal contrastive learning. In Proceedings of the 59th Annual Meeting of the Association for Computational Linguistics and the 11th International Joint Conference on Natural Language Processing (ACL\/IJCNLP\u201921). 2592\u20132607."},{"key":"e_1_3_1_27_2","article-title":"UNIMO-2: End-to-end unified vision-language grounded learning","author":"Li Wei","year":"2022","unstructured":"Wei Li, Can Gao, Guocheng Niu, Xinyan Xiao, Hao Liu, Jiachen Liu, Hua Wu, and Haifeng Wang. 2022. UNIMO-2: End-to-end unified vision-language grounded learning. In Proceedings of the Association for Computational Linguistics (ACL\u201922).","journal-title":"Proceedings of the Association for Computational Linguistics (ACL\u201922)"},{"key":"e_1_3_1_28_2","first-page":"17969","volume-title":"Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition (CVPR\u201922)","author":"Li Yehao","year":"2022","unstructured":"Yehao Li, Yingwei Pan, Ting Yao, and Tao Mei. 2022. Comprehending and ordering semantics for image captioning. In Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition (CVPR\u201922). 17969\u201317978."},{"key":"e_1_3_1_29_2","first-page":"74","volume-title":"Text Summarization Branches Out","author":"Lin Chin-Yew","year":"2004","unstructured":"Chin-Yew Lin. 2004. ROUGE: A package for automatic evaluation of summaries. In Text Summarization Branches Out. Association for Computational Linguistics, 74\u201381."},{"key":"e_1_3_1_30_2","first-page":"740","volume-title":"Proceedings of the European Conference on Computer Vision (ECCV\u201914)","volume":"8693","author":"Lin Tsung-Yi","year":"2014","unstructured":"Tsung-Yi Lin, Michael Maire, Serge J. Belongie, James Hays, Pietro Perona, Deva Ramanan, Piotr Doll\u00e1r, and C. Lawrence Zitnick. 2014. Microsoft COCO: Common objects in context. In Proceedings of the European Conference on Computer Vision (ECCV\u201914), Vol. 8693. 740\u2013755."},{"key":"e_1_3_1_31_2","first-page":"1548","volume-title":"Proceedings of the Findings of the Association for Computational Linguistics (ACL\/IJCNLP\u201921)","volume":"2021","author":"Lin Zhe","year":"2021","unstructured":"Zhe Lin and Xiaojun Wan. 2021. Pushing paraphrase away from original sentence: A multi-round paraphrase generation approach. In Proceedings of the Findings of the Association for Computational Linguistics (ACL\/IJCNLP\u201921), Chengqing Zong, Fei Xia, Wenjie Li, and Roberto Navigli (Eds.), Vol. ACL\/IJCNLP 2021. Association for Computational Linguistics, 1548\u20131557."},{"key":"e_1_3_1_32_2","first-page":"4239","volume-title":"Proceedings of the IEEE\/CVF International Conference on Computer Vision (ICCV\u201919)","author":"Liu Lixin","year":"2019","unstructured":"Lixin Liu, Jiajun Tang, Xiaojun Wan, and Zongming Guo. 2019. Generating diverse and descriptive image captions using visual paraphrases. In Proceedings of the IEEE\/CVF International Conference on Computer Vision (ICCV\u201919). 4239\u20134248."},{"key":"e_1_3_1_33_2","first-page":"4239","volume-title":"Proceedings of the IEEE\/CVF International Conference on Computer Vision (ICCV\u201919)","author":"Liu Lixin","year":"2019","unstructured":"Lixin Liu, Jiajun Tang, Xiaojun Wan, and Zongming Guo. 2019. Generating diverse and descriptive image captions using visual paraphrases. In Proceedings of the IEEE\/CVF International Conference on Computer Vision (ICCV\u201919). 4239\u20134248."},{"key":"e_1_3_1_34_2","first-page":"196","volume-title":"Proceedings of the Conference of the North American Chapter of the Association for Computational Linguistics: Human Language Technologies (NAACL-HLT\u201918)","author":"Ma Shuming","year":"2018","unstructured":"Shuming Ma, Xu Sun, Wei Li, Sujian Li, Wenjie Li, and Xuancheng Ren. 2018. Query and output: Generating words by querying distributed word representations for paraphrase generation. In Proceedings of the Conference of the North American Chapter of the Association for Computational Linguistics: Human Language Technologies (NAACL-HLT\u201918). 196\u2013206."},{"key":"e_1_3_1_35_2","volume-title":"Proceedings of the 6th International Conference on Learning Representations (ICLR\u201918)","author":"Mena Gonzalo","year":"2018","unstructured":"Gonzalo Mena, David Belanger, Scott Linderman, and Jasper Snoek. 2018. Learning latent permutations with Gumbel-Sinkhorn networks. In Proceedings of the 6th International Conference on Learning Representations (ICLR\u201918)."},{"key":"e_1_3_1_36_2","first-page":"1621","volume-title":"Proceedings of the 60th Annual Meeting of the Association for Computational Linguistics (ACL\u201922)","author":"Ormazabal Aitor","year":"2022","unstructured":"Aitor Ormazabal, Mikel Artetxe, Aitor Soroa, Gorka Labaka, and Eneko Agirre. 2022. Principled paraphrase generation with parallel corpora. In Proceedings of the 60th Annual Meeting of the Association for Computational Linguistics (ACL\u201922), Smaranda Muresan, Preslav Nakov, and Aline Villavicencio (Eds.). Association for Computational Linguistics, 1621\u20131638."},{"key":"e_1_3_1_37_2","first-page":"311","volume-title":"Proceedings of the 40th Annual Meeting of the Association for Computational Linguistics (ACL\u201902)","author":"Papineni Kishore","year":"2002","unstructured":"Kishore Papineni, Salim Roukos, Todd Ward, and Wei-Jing Zhu. 2002. Bleu: A method for automatic evaluation of machine translation. In Proceedings of the 40th Annual Meeting of the Association for Computational Linguistics (ACL\u201902). 311\u2013318."},{"key":"e_1_3_1_38_2","doi-asserted-by":"publisher","DOI":"10.1016\/j.neucom.2020.08.022"},{"key":"e_1_3_1_39_2","first-page":"2715","volume-title":"Proceedings of the 27th International Conference on Computational Linguistics (COLING\u201918)","author":"Patro Badri Narayana","year":"2018","unstructured":"Badri Narayana Patro, Vinod Kumar Kurmi, Sandeep Kumar, and Vinay P. Namboodiri. 2018. Learning semantic sentence embeddings using sequential pair-wise discriminator. In Proceedings of the 27th International Conference on Computational Linguistics (COLING\u201918). 2715\u20132729."},{"key":"e_1_3_1_40_2","first-page":"5579","volume-title":"Proceedings of the 60th Annual Meeting of the Association for Computational Linguistics (ACL\u201922)","author":"Peng Qiwei","year":"2022","unstructured":"Qiwei Peng, David J. Weir, Julie Weeds, and Yekun Chai. 2022. Predicate-argument based bi-encoder for paraphrase identification. In Proceedings of the 60th Annual Meeting of the Association for Computational Linguistics (ACL\u201922), Smaranda Muresan, Preslav Nakov, and Aline Villavicencio (Eds.). Association for Computational Linguistics, 5579\u20135589."},{"key":"e_1_3_1_41_2","doi-asserted-by":"publisher","DOI":"10.3115\/v1\/D14-1162"},{"key":"e_1_3_1_42_2","doi-asserted-by":"publisher","DOI":"10.1109\/TPAMI.2016.2577031"},{"issue":"2","key":"e_1_3_1_43_2","first-page":"54:1\u201354:20","article-title":"Show, reward, and tell: Adversarial visual story generation","volume":"15","author":"Tang Jinhui","year":"2019","unstructured":"Jinhui Tang, Jing Wang, Zechao Li, Jianlong Fu, and Tao Mei. 2019. Show, reward, and tell: Adversarial visual story generation. ACM Trans. Multim. Comput. Commun. Appl. 15, 2s (2019), 54:1\u201354:20.","journal-title":"ACM Trans. Multim. Comput. Commun. Appl."},{"key":"e_1_3_1_44_2","first-page":"4566","volume-title":"Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition (CVPR\u201915)","author":"Vedantam Ramakrishna","year":"2015","unstructured":"Ramakrishna Vedantam, C. Lawrence Zitnick, and Devi Parikh. 2015. CIDEr: Consensus-based image description evaluation. In Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition (CVPR\u201915). 4566\u20134575."},{"key":"e_1_3_1_45_2","first-page":"4052","volume-title":"Proceedings of the 60th Annual Meeting of the Association for Computational Linguistics (ACL\u201922)","author":"Weston Jack","year":"2022","unstructured":"Jack Weston, Raphael Lenain, Udeepa Meepegama, and Emil Fristed. 2022. Generative pretraining for paraphrase evaluation. In Proceedings of the 60th Annual Meeting of the Association for Computational Linguistics (ACL\u201922), Smaranda Muresan, Preslav Nakov, and Aline Villavicencio (Eds.). Association for Computational Linguistics, 4052\u20134073."},{"issue":"4","key":"e_1_3_1_46_2","doi-asserted-by":"crossref","first-page":"509","DOI":"10.2307\/358386","article-title":"A pedagogy to address plagiarism","volume":"44","author":"Whitaker Elaine E.","year":"1993","unstructured":"Elaine E. Whitaker. 1993. A pedagogy to address plagiarism. College Comp. Commun. 44, 4 (1993), 509\u2013514.","journal-title":"College Comp. Commun."},{"key":"e_1_3_1_47_2","doi-asserted-by":"publisher","DOI":"10.1609\/aaai.v31i1.10981"},{"key":"e_1_3_1_48_2","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR.2016.503"},{"key":"e_1_3_1_49_2","first-page":"79","volume-title":"Proceedings of the Conference of the North American Chapter of the Association for Computational Linguistics (NAACL\u201922)","author":"Zhao Xue","year":"2022","unstructured":"Xue Zhao, Dayiheng Liu, Junwei Ding, Liang Yao, Mahone Yan, Huibo Wang, and Wenqing Yao. 2022. Self-supervised product title rewrite for product listing ads. In Proceedings of the Conference of the North American Chapter of the Association for Computational Linguistics (NAACL\u201922), Anastassia Loukina, Rashmi Gangadharaiah, and Bonan Min (Eds.). Association for Computational Linguistics, 79\u201385."},{"key":"e_1_3_1_50_2","unstructured":"Li Zhou Jianfeng Gao Di Li and Heung-Yeung Shum. 2018. The design and implementation of XiaoIce an empathetic social chatbot. Retrieved from https:\/\/abs\/1812.08989."},{"key":"e_1_3_1_51_2","doi-asserted-by":"crossref","first-page":"1","DOI":"10.1145\/3366710","article-title":"Multichannel attention refinement for video question answering","volume":"16","author":"Zhuang Yueting","year":"2020","unstructured":"Yueting Zhuang, Dejing Xu, Xin Yan, Wenzhuo Cheng, Zhou Zhao, Shiliang Pu, and Jun Xiao. 2020. Multichannel attention refinement for video question answering. ACM Trans. Multim. Comput. Commun. Appl. 16, Special (2020), 1\u201323.","journal-title":"ACM Trans. Multim. Comput. Commun. Appl."}],"container-title":["ACM Transactions on Multimedia Computing, Communications, and Applications"],"original-title":[],"language":"en","link":[{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/3585010","content-type":"unspecified","content-version":"vor","intended-application":"text-mining"},{"URL":"https:\/\/dl.acm.org\/doi\/pdf\/10.1145\/3585010","content-type":"unspecified","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2025,6,17]],"date-time":"2025-06-17T16:37:07Z","timestamp":1750178227000},"score":1,"resource":{"primary":{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/3585010"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2023,5,30]]},"references-count":50,"journal-issue":{"issue":"6","published-print":{"date-parts":[[2023,11,30]]}},"alternative-id":["10.1145\/3585010"],"URL":"https:\/\/doi.org\/10.1145\/3585010","relation":{},"ISSN":["1551-6857","1551-6865"],"issn-type":[{"type":"print","value":"1551-6857"},{"type":"electronic","value":"1551-6865"}],"subject":[],"published":{"date-parts":[[2023,5,30]]},"assertion":[{"value":"2022-06-21","order":0,"name":"received","label":"Received","group":{"name":"publication_history","label":"Publication History"}},{"value":"2022-12-24","order":1,"name":"accepted","label":"Accepted","group":{"name":"publication_history","label":"Publication History"}},{"value":"2023-05-30","order":2,"name":"published","label":"Published","group":{"name":"publication_history","label":"Publication History"}}]}}