{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2025,6,19]],"date-time":"2025-06-19T05:02:17Z","timestamp":1750309337176,"version":"3.41.0"},"reference-count":34,"publisher":"Association for Computing Machinery (ACM)","issue":"3","license":[{"start":{"date-parts":[[2024,3,9]],"date-time":"2024-03-09T00:00:00Z","timestamp":1709942400000},"content-version":"vor","delay-in-days":0,"URL":"https:\/\/www.acm.org\/publications\/policies\/copyright_policy#Background"}],"funder":[{"name":"Scientific and Technological Innovation 2030 - major project of new generation artificial intelligence","award":["2020AAA0109300"],"award-info":[{"award-number":["2020AAA0109300"]}]}],"content-domain":{"domain":["dl.acm.org"],"crossmark-restriction":true},"short-container-title":["ACM Trans. Asian Low-Resour. Lang. Inf. Process."],"published-print":{"date-parts":[[2024,3,31]]},"abstract":"<jats:p>With its unique information-filtering function, text summarization technology has become a significant aspect of search engines and question-and-answer systems. However, existing models that include the copy mechanism often lack the ability to extract important fragments, resulting in generated content that suffers from thematic deviation and insufficient generalization. Specifically, Chinese automatic summarization using traditional generation methods often loses semantics because of its reliance on word lists. To address these issues, we proposed the novel BioCopy mechanism for the summarization task. By training the tags of predictive words and reducing the probability distribution range on the glossary, we enhanced the ability to generate continuous segments, which effectively solves the above problems. Additionally, we applied reinforced canonicality to the inputs to obtain better model results, making the model share the sub-network weight parameters and sparsing the model output to reduce the search space for model prediction. To further improve the model\u2019s performance, we calculated the bilingual evaluation understudy (BLEU) score on the English dataset CNN\/DailyMail to filter the thresholds and reduce the difficulty of word separation and the dependence of the output on the word list. We fully fine-tuned the model using the LCSTS dataset for the Chinese summarization task and conducted small-sample experiments using the CSL dataset. We also conducted ablation experiments on the Chinese dataset. The experimental results demonstrate that the optimized model can learn the semantic representation of the original text better than other models and performs well with small sample sizes.<\/jats:p>","DOI":"10.1145\/3643695","type":"journal-article","created":{"date-parts":[[2024,2,6]],"date-time":"2024-02-06T07:02:50Z","timestamp":1707202970000},"page":"1-16","update-policy":"https:\/\/doi.org\/10.1145\/crossmark-policy","source":"Crossref","is-referenced-by-count":0,"title":["Improved BIO-Based Chinese Automatic Abstract-Generation Model"],"prefix":"10.1145","volume":"23","author":[{"ORCID":"https:\/\/orcid.org\/0000-0002-2682-6818","authenticated-orcid":false,"given":"Qing","family":"Li","sequence":"first","affiliation":[{"name":"Shanghai University of Engineering Science, Songjiang Qu, China"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"ORCID":"https:\/\/orcid.org\/0000-0002-7092-9849","authenticated-orcid":false,"given":"Weibin","family":"Wan","sequence":"additional","affiliation":[{"name":"Shanghai University of Engineering Science, Songjiang Qu, China"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"ORCID":"https:\/\/orcid.org\/0009-0007-1443-3463","authenticated-orcid":false,"given":"Yuming","family":"Zhao","sequence":"additional","affiliation":[{"name":"Shanghai Jiao Tong University, Minhang Qu, China"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"ORCID":"https:\/\/orcid.org\/0000-0002-1946-576X","authenticated-orcid":false,"given":"Xiaoyan","family":"Jiang","sequence":"additional","affiliation":[{"name":"Shanghai University of Engineering Science, Songjiang Qu, China"}],"role":[{"role":"author","vocabulary":"crossref"}]}],"member":"320","published-online":{"date-parts":[[2024,3,9]]},"reference":[{"key":"e_1_3_1_2_2","doi-asserted-by":"publisher","DOI":"10.17261\/Pressacademia.2017.591"},{"key":"e_1_3_1_3_2","doi-asserted-by":"publisher","DOI":"10.1145\/3314934"},{"key":"e_1_3_1_4_2","article-title":"Text summarization techniques: A brief survey","author":"Allahyari Mehdi","year":"2017","unstructured":"Mehdi Allahyari, Seyedamin Pouriyeh, Mehdi Assefi, Saeid Safaei, Elizabeth D. Trippe, Juan B. Gutierrez, and Krys Kochut. 2017. Text summarization techniques: A brief survey. arXiv preprint arXiv:1707.02268 (2017).","journal-title":"arXiv preprint arXiv:1707.02268"},{"key":"e_1_3_1_5_2","first-page":"5998","volume-title":"Advances in Neural Information Processing Systems","author":"Vaswani Ashish","year":"2017","unstructured":"Ashish Vaswani, Noam Shazeer, Niki Parmar, Jakob Uszkoreit, Llion Jones, Aidan N. Gomez, \u0141ukasz Kaiser, and Illia Polosukhin. 2017. Attention is all you need. In Advances in Neural Information Processing Systems. 5998\u20136008."},{"key":"e_1_3_1_6_2","first-page":"1","article-title":"Pre-trained models for natural language processing: A survey","author":"Qiu Xipeng","year":"2020","unstructured":"Xipeng Qiu, Tianxiang Sun, Yige Xu, Yunfan Shao, Ning Dai, and Xuanjing Huang. 2020. Pre-trained models for natural language processing: A survey. Science China Technological Sciences (2020), 1\u201326.","journal-title":"Science China Technological Sciences"},{"key":"e_1_3_1_7_2","article-title":"mt5: A massively multilingual pre-trained text-to-text transformer","author":"Xue Linting","year":"2020","unstructured":"Linting Xue, Noah Constant, Adam Roberts, Mihir Kale, Rami Al-Rfou, Aditya Siddhant, Aditya Barua, and Colin Raffel. 2020. mt5: A massively multilingual pre-trained text-to-text transformer. arXiv preprint arXiv:2010.11934 (2020).","journal-title":"arXiv preprint arXiv:2010.11934"},{"key":"e_1_3_1_8_2","article-title":"R-drop: Regularized dropout for neural networks","volume":"34","author":"Wu Lijun","year":"2021","unstructured":"Lijun Wu, Juntao Li, Yue Wang, Qi Meng, Tao Qin, Wei Chen, Min Zhang, and Tie-Yan Liu. 2021. R-drop: Regularized dropout for neural networks. Advances in Neural Information Processing Systems 34 (2021).","journal-title":"Advances in Neural Information Processing Systems"},{"key":"e_1_3_1_9_2","first-page":"1614","volume-title":"International Conference on Machine Learning","author":"Martins Andre","year":"2016","unstructured":"Andre Martins and Ramon Astudillo. 2016. From softmax to sparsemax: A sparse model of attention and multi-label classification. In International Conference on Machine Learning. PMLR, 1614\u20131623."},{"key":"e_1_3_1_10_2","article-title":"BioCopy: A plug-and-play span copy mechanism in Seq2Seq models","author":"Liu Yi","year":"2021","unstructured":"Yi Liu, Guoan Zhang, Puning Yu, Jianlin Su, and Shengfeng Pan. 2021. BioCopy: A plug-and-play span copy mechanism in Seq2Seq models. arXiv preprint arXiv:2109.12533 (2021).","journal-title":"arXiv preprint arXiv:2109.12533"},{"key":"e_1_3_1_11_2","first-page":"3104","volume-title":"Advances in Neural Information Processing Systems","author":"Sutskever Ilya","year":"2014","unstructured":"Ilya Sutskever, Oriol Vinyals, and Quoc V. Le. 2014. Sequence to sequence learning with neural networks. In Advances in Neural Information Processing Systems. 3104\u20133112."},{"key":"e_1_3_1_12_2","article-title":"A neural attention model for abstractive sentence summarization","author":"Rush Alexander M.","year":"2015","unstructured":"Alexander M. Rush, Sumit Chopra, and Jason Weston. 2015. A neural attention model for abstractive sentence summarization. arXiv preprint arXiv:1509.00685 (2015).","journal-title":"arXiv preprint arXiv:1509.00685"},{"key":"e_1_3_1_13_2","doi-asserted-by":"publisher","DOI":"10.18653\/v1\/N16-1012"},{"key":"e_1_3_1_14_2","doi-asserted-by":"publisher","DOI":"10.1609\/aaai.v31i1.10958"},{"key":"e_1_3_1_15_2","article-title":"Get to the point: Summarization with pointer-generator networks","author":"See Abigail","year":"2017","unstructured":"Abigail See, Peter J. Liu, and Christopher D. Manning. 2017. Get to the point: Summarization with pointer-generator networks. arXiv preprint arXiv:1704.04368 (2017).","journal-title":"arXiv preprint arXiv:1704.04368"},{"key":"e_1_3_1_16_2","article-title":"Incorporating copying mechanism in sequence-to-sequence learning","author":"Gu Jiatao","year":"2016","unstructured":"Jiatao Gu, Zhengdong Lu, Hang Li, and Victor O. K. Li. 2016. Incorporating copying mechanism in sequence-to-sequence learning. arXiv preprint arXiv:1603.06393 (2016).","journal-title":"arXiv preprint arXiv:1603.06393"},{"key":"e_1_3_1_17_2","article-title":"Efficient estimation of word representations in vector space","author":"Mikolov Tomas","year":"2013","unstructured":"Tomas Mikolov, Kai Chen, Greg Corrado, and Jeffrey Dean. 2013. Efficient estimation of word representations in vector space. arXiv preprint arXiv:1301.3781 (2013).","journal-title":"arXiv preprint arXiv:1301.3781"},{"key":"e_1_3_1_18_2","doi-asserted-by":"crossref","unstructured":"Jeffrey Pennington Richard Socher and Christopher Manning. 2014. GloVe: global vectors for word representation. In Proceedings of the 2014 Conference on Empirical Methods in Natural Language Processing (EMNLP) Doha Qatar. Association for Computational Linguistics 532\u20131543.","DOI":"10.3115\/v1\/D14-1162"},{"key":"e_1_3_1_19_2","article-title":"Deep contextualized word representations","volume":"1802","author":"Peters Matthew E.","year":"2018","unstructured":"Matthew E. Peters, Mark Neumann, Mohit Iyyer, Matt Gardner, Christopher Clark, Kenton Lee, and Luke Zettlemoyer. 2018. Deep contextualized word representations. CoRR abs\/1802.05365 (2018). arXiv:1802.05365http:\/\/arxiv.org\/abs\/1802.05365","journal-title":"CoRR"},{"key":"e_1_3_1_20_2","doi-asserted-by":"publisher","DOI":"10.1007\/978-3-030-32381-3_16"},{"key":"e_1_3_1_21_2","unstructured":"Alec Radford Karthik Narasimhan Tim Salimans and Ilya Sutskever. 2018. Improving language understanding by generative pre-training. (2018)."},{"key":"e_1_3_1_22_2","article-title":"Self-attention with relative position representations","author":"Shaw Peter","year":"2018","unstructured":"Peter Shaw, Jakob Uszkoreit, and Ashish Vaswani. 2018. Self-attention with relative position representations. arXiv preprint arXiv:1803.02155 (2018).","journal-title":"arXiv preprint arXiv:1803.02155"},{"key":"e_1_3_1_23_2","first-page":"933","volume-title":"International Conference on Machine Learning","author":"Dauphin Yann N.","year":"2017","unstructured":"Yann N. Dauphin, Angela Fan, Michael Auli, and David Grangier. 2017. Language modeling with gated convolutional networks. In International Conference on Machine Learning. PMLR, 933\u2013941."},{"key":"e_1_3_1_24_2","article-title":"GLU variants improve transformer","volume":"2002","author":"Shazeer Noam","year":"2020","unstructured":"Noam Shazeer. 2020. GLU variants improve transformer. CoRR abs\/2002.05202 (2020). arXiv:2002.05202https:\/\/arxiv.org\/abs\/2002.05202","journal-title":"CoRR"},{"key":"e_1_3_1_25_2","article-title":"Sparse sequence-to-sequence models","author":"Peters Ben","year":"2019","unstructured":"Ben Peters, Vlad Niculae, and Andr\u00e9 F. T. Martins. 2019. Sparse sequence-to-sequence models. arXiv preprint arXiv:1905.05702 (2019).","journal-title":"arXiv preprint arXiv:1905.05702"},{"key":"e_1_3_1_26_2","article-title":"How multilingual is multilingual BERT?","author":"Pires Telmo","year":"2019","unstructured":"Telmo Pires, Eva Schlinger, and Dan Garrette. 2019. How multilingual is multilingual BERT? arXiv preprint arXiv:1906.01502 (2019).","journal-title":"arXiv preprint arXiv:1906.01502"},{"key":"e_1_3_1_27_2","first-page":"74","volume-title":"Text Summarization Branches Out","author":"Lin Chin-Yew","year":"2004","unstructured":"Chin-Yew Lin. 2004. ROUGE: A package for automatic evaluation of summaries. In Text Summarization Branches Out. 74\u201381."},{"key":"e_1_3_1_28_2","first-page":"311","volume-title":"40th Annual Meeting of the Association for Computational Linguistics","author":"Papineni Kishore","year":"2002","unstructured":"Kishore Papineni, Salim Roukos, Todd Ward, and Wei-Jing Zhu. 2002. BLEU: A method for automatic evaluation of machine translation. In 40th Annual Meeting of the Association for Computational Linguistics. 311\u2013318."},{"key":"e_1_3_1_29_2","first-page":"1693","article-title":"Teaching machines to read and comprehend","volume":"28","author":"Hermann Karl Moritz","year":"2015","unstructured":"Karl Moritz Hermann, Tomas Kocisky, Edward Grefenstette, Lasse Espeholt, Will Kay, Mustafa Suleyman, and Phil Blunsom. 2015. Teaching machines to read and comprehend. Advances in Neural Information Processing Systems 28 (2015), 1693\u20131701.","journal-title":"Advances in Neural Information Processing Systems"},{"key":"e_1_3_1_30_2","article-title":"Abstractive text summarization using sequence-to-sequence rnns and beyond","author":"Nallapati Ramesh","year":"2016","unstructured":"Ramesh Nallapati, Bowen Zhou, Caglar Gulcehre, and Bing Xiang. 2016. Abstractive text summarization using sequence-to-sequence rnns and beyond. arXiv preprint arXiv:1602.06023 (2016).","journal-title":"arXiv preprint arXiv:1602.06023"},{"key":"e_1_3_1_31_2","article-title":"LCSTS: A large-scale Chinese short text summarization dataset","author":"Hu Baotian","year":"2015","unstructured":"Baotian Hu, Qingcai Chen, and Fangze Zhu. 2015. LCSTS: A large-scale Chinese short text summarization dataset. arXiv preprint arXiv:1506.05865 (2015).","journal-title":"arXiv preprint arXiv:1506.05865"},{"key":"e_1_3_1_32_2","doi-asserted-by":"publisher","DOI":"10.1109\/BigMM.2016.44"},{"key":"e_1_3_1_33_2","article-title":"Load what you need: Smaller versions of multilingual BERT","author":"Abdaoui Amine","year":"2020","unstructured":"Amine Abdaoui, Camille Pradel, and Gr\u00e9goire Sigel. 2020. Load what you need: Smaller versions of multilingual BERT. arXiv preprint arXiv:2010.05609 (2020).","journal-title":"arXiv preprint arXiv:2010.05609"},{"key":"e_1_3_1_34_2","first-page":"4596","volume-title":"International Conference on Machine Learning","author":"Shazeer Noam","year":"2018","unstructured":"Noam Shazeer and Mitchell Stern. 2018. Adafactor: Adaptive learning rates with sublinear memory cost. In International Conference on Machine Learning. PMLR, 4596\u20134604."},{"key":"e_1_3_1_35_2","article-title":"Unified language model pre-training for natural language understanding and generation","volume":"32","author":"Dong Li","year":"2019","unstructured":"Li Dong, Nan Yang, Wenhui Wang, Furu Wei, Xiaodong Liu, Yu Wang, Jianfeng Gao, Ming Zhou, and Hsiao-Wuen Hon. 2019. Unified language model pre-training for natural language understanding and generation. Advances in Neural Information Processing Systems 32 (2019).","journal-title":"Advances in Neural Information Processing Systems"}],"container-title":["ACM Transactions on Asian and Low-Resource Language Information Processing"],"original-title":[],"language":"en","link":[{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/3643695","content-type":"unspecified","content-version":"vor","intended-application":"text-mining"},{"URL":"https:\/\/dl.acm.org\/doi\/pdf\/10.1145\/3643695","content-type":"unspecified","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2025,6,19]],"date-time":"2025-06-19T00:04:20Z","timestamp":1750291460000},"score":1,"resource":{"primary":{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/3643695"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2024,3,9]]},"references-count":34,"journal-issue":{"issue":"3","published-print":{"date-parts":[[2024,3,31]]}},"alternative-id":["10.1145\/3643695"],"URL":"https:\/\/doi.org\/10.1145\/3643695","relation":{},"ISSN":["2375-4699","2375-4702"],"issn-type":[{"type":"print","value":"2375-4699"},{"type":"electronic","value":"2375-4702"}],"subject":[],"published":{"date-parts":[[2024,3,9]]},"assertion":[{"value":"2022-01-31","order":0,"name":"received","label":"Received","group":{"name":"publication_history","label":"Publication History"}},{"value":"2024-01-23","order":1,"name":"accepted","label":"Accepted","group":{"name":"publication_history","label":"Publication History"}},{"value":"2024-03-09","order":2,"name":"published","label":"Published","group":{"name":"publication_history","label":"Publication History"}}]}}