{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2025,11,24]],"date-time":"2025-11-24T12:50:11Z","timestamp":1763988611321,"version":"3.45.0"},"reference-count":70,"publisher":"Springer Science and Business Media LLC","issue":"10","license":[{"start":{"date-parts":[[2025,7,24]],"date-time":"2025-07-24T00:00:00Z","timestamp":1753315200000},"content-version":"tdm","delay-in-days":0,"URL":"https:\/\/creativecommons.org\/licenses\/by\/4.0"},{"start":{"date-parts":[[2025,7,24]],"date-time":"2025-07-24T00:00:00Z","timestamp":1753315200000},"content-version":"vor","delay-in-days":0,"URL":"https:\/\/creativecommons.org\/licenses\/by\/4.0"}],"funder":[{"DOI":"10.13039\/501100013076","name":"National Major Science and Technology Projects of China","doi-asserted-by":"publisher","award":["2022ZD0116204"],"award-info":[{"award-number":["2022ZD0116204"]}],"id":[{"id":"10.13039\/501100013076","id-type":"DOI","asserted-by":"publisher"}]},{"DOI":"10.13039\/501100012456","name":"National Social Science Fund of China","doi-asserted-by":"publisher","award":["23BTQ081"],"award-info":[{"award-number":["23BTQ081"]}],"id":[{"id":"10.13039\/501100012456","id-type":"DOI","asserted-by":"publisher"}]},{"DOI":"10.13039\/501100001809","name":"National Natural Science Foundation of China","doi-asserted-by":"publisher","award":["72234005"],"award-info":[{"award-number":["72234005"]}],"id":[{"id":"10.13039\/501100001809","id-type":"DOI","asserted-by":"publisher"}]}],"content-domain":{"domain":["link.springer.com"],"crossmark-restriction":false},"short-container-title":["Scientometrics"],"published-print":{"date-parts":[[2025,10]]},"abstract":"<jats:title>Abstract<\/jats:title>\n                  <jats:p>\n                    The automatic recognition of functional structures within scientific literature enhances fine-grained information retrieval and mitigates issues associated with imbalanced text classification. Although multilevel functional structure research is relatively advanced, achieving high accuracy in overall label prediction and recognition, precise functional structure identification at the paragraph level remains challenging. To address this issue, we propose an innovative method called\n                    <jats:bold>SLSG<\/jats:bold>\n                    , which stands for (\n                    <jats:bold>S<\/jats:bold>\n                    ynonym replacement +\n                    <jats:bold>L<\/jats:bold>\n                    exical function based LLM Auto-labeling +\n                    <jats:bold>S<\/jats:bold>\n                    ciBERT-\n                    <jats:bold>G<\/jats:bold>\n                    CN). This method serves as a data augmentation (DA) strategy for the identification of paragraph-level functional structure. Specifically,\n                    <jats:bold>SLSG<\/jats:bold>\n                    integrates several mechanisms, including synonym replacement and lexical function-based auto-annotation using Large Language Models (LLMs) for DA. It combines augmented data with a SciBERT-GCN model to effectively extract features by leveraging contextual sequence information between paragraphs. Applied to the ScienceDirect dataset,\n                    <jats:bold>SLSG<\/jats:bold>\n                    achieves an F1 score of 86% for paragraph-level functional structure recognition, marking an 18% improvement over the baseline models and demonstrating a significant enhancement in classification performance. Moreover, SLSG employs graph neural networks to capture both dependency relationships and topological structures among word nodes. This approach not only augments the representation of scientific literature, but establishes a solid research paradigm to address the challenges related to unbalanced text classification.\n                  <\/jats:p>","DOI":"10.1007\/s11192-025-05355-6","type":"journal-article","created":{"date-parts":[[2025,7,24]],"date-time":"2025-07-24T16:06:10Z","timestamp":1753373170000},"page":"5473-5502","update-policy":"https:\/\/doi.org\/10.1007\/springer_crossmark_policy","source":"Crossref","is-referenced-by-count":1,"title":["Research on paragraph-level functional structure recognition in scientific literature: a data augmentation method based on LLMs and lexical function"],"prefix":"10.1007","volume":"130","author":[{"given":"Haotan","family":"Liu","sequence":"first","affiliation":[],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Zhuo","family":"Chen","sequence":"additional","affiliation":[],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Qunzhe","family":"Ding","sequence":"additional","affiliation":[],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Jiafeng","family":"Zhang","sequence":"additional","affiliation":[],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Jiawei","family":"Liu","sequence":"additional","affiliation":[],"role":[{"role":"author","vocabulary":"crossref"}]},{"ORCID":"https:\/\/orcid.org\/0000-0001-6491-1995","authenticated-orcid":false,"given":"Jiming","family":"Hu","sequence":"additional","affiliation":[],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Wei","family":"Lu","sequence":"additional","affiliation":[],"role":[{"role":"author","vocabulary":"crossref"}]}],"member":"297","published-online":{"date-parts":[[2025,7,24]]},"reference":[{"key":"5355_CR1","unstructured":"Alizadeh, M., Kubli, M., Samei, Z., Dehghani, S., Bermeo, J.D., Korobeynikova, M., & Gilardi, F.: (2023). Open-source large language models outperform crowd workers and approach chatgpt in text-annotation tasks. arXiv preprint 101 arXiv:2307.02179."},{"issue":"4","key":"5355_CR2","doi-asserted-by":"publisher","first-page":"1011","DOI":"10.1016\/j.jksuci.2020.04.020","volume":"34","author":"NI Altmami","year":"2022","unstructured":"Altmami, N. I., & Menai, M. E. B. (2022). Automatic summarization of scientific articles: A survey. Journal of King Saud University-Computer and Information Sciences, 34(4), 1011\u20131028.","journal-title":"Journal of King Saud University-Computer and Information Sciences"},{"key":"5355_CR3","unstructured":"Bansal, P., & Sharma, A. (2023). Large language models as annotators: Enhancing generalization of nlp models at minimal cost. arXiv preprint arXiv:2306.15766."},{"key":"5355_CR4","doi-asserted-by":"publisher","first-page":"43","DOI":"10.1016\/j.eswa.2016.01.007","volume":"53","author":"R Belkebir","year":"2016","unstructured":"Belkebir, R., & Guessoum, A. (2016). Concept generalization and fusion for abstractive sentence generation. Expert Systems with Applications, 53, 43\u201356.","journal-title":"Expert Systems with Applications"},{"key":"5355_CR5","doi-asserted-by":"crossref","unstructured":"Bertaglia, T., Huber, S., Goanta, C., Spanakis, G., & Iamnitchi, A. (2023). Closing the loop: Testing chatgpt to generate model explanations to improve human labelling of sponsored content on social media. In: World Conference on Explainable Artificial Intelligence, pp. 198\u2013213 Springer.","DOI":"10.1007\/978-3-031-44067-0_11"},{"issue":"1","key":"5355_CR6","doi-asserted-by":"publisher","first-page":"237","DOI":"10.1016\/j.hrmr.2016.09.013","volume":"27","author":"FA Bosco","year":"2017","unstructured":"Bosco, F. A., Uggerslev, K. L., & Steel, P. (2017). Metabus as a vehicle for facilitating meta-analysis. Human Resource Management Review, 27(1), 237\u2013254.","journal-title":"Human Resource Management Review"},{"issue":"2","key":"5355_CR7","first-page":"1488","volume":"13","author":"W Cao","year":"2023","unstructured":"Cao, W., Wu, Y., Sun, Y., Zhang, H., Ren, J., Gu, D., & Wang, X. (2023). A review on multimodal zero-shot learning. Wiley Interdisciplinary Reviews: Data Mining and Knowledge Discovery, 13(2), 1488.","journal-title":"Wiley Interdisciplinary Reviews: Data Mining and Knowledge Discovery"},{"key":"5355_CR8","doi-asserted-by":"crossref","unstructured":"Chen, Z., Gao, Q., Bosselut, A., Sabharwal, A., & Richardson, K. (2022). Disco: Distilling counterfactuals with large language models. arXiv preprint arXiv:2212.10534.","DOI":"10.18653\/v1\/2023.acl-long.302"},{"key":"5355_CR9","doi-asserted-by":"publisher","first-page":"191","DOI":"10.1162\/tacl_a_00542","volume":"11","author":"J Chen","year":"2023","unstructured":"Chen, J., Tam, D., Raffel, C., Bansal, M., & Yang, D. (2023). An empirical survey of data augmentation for limited data learning in nlp. Transactions of the Association for Computational Linguistics, 11, 191\u2013211.","journal-title":"Transactions of the Association for Computational Linguistics"},{"key":"5355_CR10","doi-asserted-by":"crossref","unstructured":"Chintagunta, B., Katariya, N., Amatriain, X., & Kannan, A. (2021). Medically aware gpt-3 as a data generator for medical dialogue summarization. In: Machine Learning for Healthcare Conference, pp. 354\u2013372 PMLR.","DOI":"10.18653\/v1\/2021.nlpmc-1.9"},{"key":"5355_CR11","unstructured":"Dai, H., Liu, Z., Liao, W., Huang, X., Cao, Y., Wu, Z., Zhao, L., Xu, S., Liu, W., & Liu, N., et al.: (2023). Auggpt: Leveraging chatgpt for text data augmentation. arXiv preprint arXiv:2302.13007."},{"key":"5355_CR12","doi-asserted-by":"crossref","unstructured":"Ding, B., Qin, C., Zhao, R., Luo, T., Li, X., Chen, G., Xia, W., Hu, J., Tuan, L.A., & Joty, S. (2024). Data augmentation using llms: Data perspectives, learning paradigms and challenges. In: Findings of the Association for Computational Linguistics ACL 2024, pp. 1679\u20131705.","DOI":"10.18653\/v1\/2024.findings-acl.97"},{"key":"5355_CR13","doi-asserted-by":"crossref","unstructured":"Dixit, T., Paranjape, B., Hajishirzi, H., & Zettlemoyer, L. (2022). Core: A retrieve-then-edit framework for counterfactual data generation. arXiv preprint arXiv:2210.04873.","DOI":"10.18653\/v1\/2022.findings-emnlp.216"},{"key":"5355_CR14","doi-asserted-by":"crossref","unstructured":"Edunov, S. (2018). Understanding back-translation at scale. arXiv preprint arXiv:1808.09381.","DOI":"10.18653\/v1\/D18-1045"},{"key":"5355_CR15","doi-asserted-by":"crossref","unstructured":"Gao, L., Liu, W., Liu, K., & Wu, J. (2024). Augsteal: Advancing model steal with data augmentation in active learning frameworks. IEEE Transactions on Information Forensics and Security.","DOI":"10.1109\/TIFS.2024.3384841"},{"key":"5355_CR16","doi-asserted-by":"publisher","DOI":"10.1016\/j.est.2023.107266","volume":"66","author":"M Ghalambaz","year":"2023","unstructured":"Ghalambaz, M., Sheremet, M., Fauzi, M. A., Fteiti, M., & Younis, O. (2023). A scientometrics review of solar thermal energy storage (stes) during the past forty years. Journal of Energy Storage, 66, Article 107266.","journal-title":"Journal of Energy Storage"},{"issue":"30","key":"5355_CR17","doi-asserted-by":"publisher","first-page":"2305016120","DOI":"10.1073\/pnas.2305016120","volume":"120","author":"F Gilardi","year":"2023","unstructured":"Gilardi, F., Alizadeh, M., & Kubli, M. (2023). Chatgpt outperforms crowd workers for text-annotation tasks. Proceedings of the National Academy of Sciences, 120(30), 2305016120.","journal-title":"Proceedings of the National Academy of Sciences"},{"key":"5355_CR18","unstructured":"Grattafiori, A., Dubey, A., Jauhri, A., Pandey, A., Kadian, A., Al-Dahle, A., Letman, A., Mathur, A., Schelten, A., & Vaughan, A., et al. (2024). The llama 3 herd of models. arXiv preprint arXiv:2407.21783."},{"issue":"10","key":"5355_CR19","doi-asserted-by":"publisher","first-page":"2181","DOI":"10.1093\/jamia\/ocae210","volume":"31","author":"Y Guo","year":"2024","unstructured":"Guo, Y., Ovadje, A., Al-Garadi, M. A., & Sarker, A. (2024). Evaluating large language models for health-related text classification tasks with public social media data. Journal of the American Medical Informatics Association, 31(10), 2181\u20132189.","journal-title":"Journal of the American Medical Informatics Association"},{"key":"5355_CR20","doi-asserted-by":"crossref","unstructured":"He, Q., Pei, J., Kifer, D., Mitra, P., & Giles, L. (2010). Context-aware citation recommendation. In: Proceedings of the 19th International Conference on World Wide Web, pp. 421\u2013430","DOI":"10.1145\/1772690.1772734"},{"key":"5355_CR21","doi-asserted-by":"crossref","unstructured":"Hu, Z., Dong, Y., Wang, K., & Sun, Y. (2020). Heterogeneous graph transformer. In: Proceedings of the Web Conference 2020, pp. 2704\u20132710.","DOI":"10.1145\/3366423.3380027"},{"issue":"03","key":"5355_CR22","first-page":"293","volume":"35","author":"Y Huang","year":"2016","unstructured":"Huang, Y., Lu, W., & Cheng, Q. (2016). The structure recognition of academic text chapter content based recognition. Journal of the China Society for Scientific and Technical Information, 35(03), 293\u2013300.","journal-title":"Journal of the China Society for Scientific and Technical Information"},{"issue":"4","key":"5355_CR23","first-page":"425","volume":"35","author":"Y Huang","year":"2016","unstructured":"Huang, Y., Lu, W., Cheng, Q., & Gui, S. (2016). The structure function recognition of academic text\u2014application in academic search. Journal of the China Society for Scientific and Technical Information, 35(5), 530\u2013538.","journal-title":"Journal of the China Society for Scientific and Technical Information"},{"issue":"7","key":"5355_CR24","doi-asserted-by":"publisher","first-page":"1025","DOI":"10.1002\/asi.24610","volume":"73","author":"S Huang","year":"2022","unstructured":"Huang, S., Qian, J., Huang, Y., Lu, W., Bu, Y., Yang, J., & Cheng, Q. (2022). Disclosing the relationship between citation structure and future impact of a publication. Journal of the Association for Information Science and Technology, 73(7), 1025\u20131042.","journal-title":"Journal of the Association for Information Science and Technology"},{"issue":"2","key":"5355_CR25","first-page":"152","volume":"40","author":"Y Jiang","year":"2021","unstructured":"Jiang, Y., Huang, Y., Xia, Y., Li, P., & Lu, W. (2021). Recognition of lexical functions in academic texts: Application in automatic keyword extraction. Journal of the China Society for Scientific and Technical Information, 40(2), 152\u2013162.","journal-title":"Journal of the China Society for Scientific and Technical Information"},{"key":"5355_CR26","doi-asserted-by":"crossref","unstructured":"Jin, L., & Gildea, D. (2022). Rewarding semantic similarity under optimized alignments for amr-to-text generation. In: Proceedings of the 60th Annual Meeting of the Association for Computational Linguistics (Volume 2: Short Papers), pp. 710\u2013715.","DOI":"10.18653\/v1\/2022.acl-short.80"},{"issue":"10","key":"5355_CR27","first-page":"3806","volume":"33","author":"XM Kang","year":"2022","unstructured":"Kang, X. M., & Zong, C. Q. (2022). Neural machine translation based on multi-task learning of discourse structure. Journal of Software, 33(10), 3806\u20133818.","journal-title":"Journal of Software"},{"key":"5355_CR28","doi-asserted-by":"publisher","DOI":"10.1016\/j.eswa.2023.121199","volume":"235","author":"M Kuntalp","year":"2024","unstructured":"Kuntalp, M., & D\u00fczyel, O. (2024). A new method for gan-based data augmentation for classes with distinct clusters. Expert Systems with Applications, 235, Article 121199.","journal-title":"Expert Systems with Applications"},{"key":"5355_CR29","doi-asserted-by":"crossref","unstructured":"Labruna, T., Brenna, S., Zaninello, A., & Magnini, B. (2023). Unraveling chatgpt: A critical analysis of ai-generated goal-oriented dialogues and annotations. In: International Conference of the Italian Association for Artificial Intelligence, pp. 151\u2013171 Springer.","DOI":"10.1007\/978-3-031-47546-7_11"},{"key":"5355_CR30","unstructured":"Latif, S., Usama, M., Malik, M.I., & Schuller, B.W. (2023). Can large language models aid in annotating speech emotional data? uncovering new frontiers. arXiv preprint arXiv:2307.06090."},{"issue":"4","key":"5355_CR31","doi-asserted-by":"publisher","first-page":"1234","DOI":"10.1093\/bioinformatics\/btz682","volume":"36","author":"J Lee","year":"2020","unstructured":"Lee, J., Yoon, W., Kim, S., Kim, D., Kim, S., So, C. H., & Kang, J. (2020). Biobert: A pre-trained biomedical language representation model for biomedical text mining. Bioinformatics, 36(4), 1234\u20131240.","journal-title":"Bioinformatics"},{"key":"5355_CR32","doi-asserted-by":"crossref","unstructured":"Li, Z., Chen, W., Li, S., Wang, H., Qian, J., & Yan, X. (2022). Controllable dialogue simulation with in-context learning. arXiv preprint arXiv:2210.04185","DOI":"10.18653\/v1\/2022.findings-emnlp.318"},{"key":"5355_CR33","doi-asserted-by":"crossref","unstructured":"Li, G., Muller, M., Thabet, A., & Ghanem, B. (2019). Deepgcns: Can gcns go as deep as cnns? In: Proceedings of the IEEE\/CVF International Conference on Computer Vision, pp. 9267\u20139276.","DOI":"10.1109\/ICCV.2019.00936"},{"key":"5355_CR34","doi-asserted-by":"crossref","unstructured":"Li, M., Shi, T., Ziems, C., Kan, M.-Y., Chen, N.F., Liu, Z., & Yang, D. (2023). Coannotating: Uncertainty-guided work allocation between human and large language models for data annotation. arXiv preprint arXiv:2310.15638.","DOI":"10.18653\/v1\/2023.emnlp-main.92"},{"key":"5355_CR35","doi-asserted-by":"crossref","unstructured":"Li, C., Su, Y., & Liu, W. (2018). Text-to-text generative adversarial networks. In: 2018 International Joint Conference on Neural Networks (IJCNN), pp. 1\u20137 IEEE","DOI":"10.1109\/IJCNN.2018.8489624"},{"key":"5355_CR36","doi-asserted-by":"crossref","unstructured":"Lin, Y., Meng, Y., Sun, X., Han, Q., Kuang, K., Li, J., & Wu, F. (2021). Bertgcn: Transductive text classification by combining gcn and bert. arXiv preprint arXiv:2105.05727.","DOI":"10.18653\/v1\/2021.findings-acl.126"},{"issue":"3","key":"5355_CR37","first-page":"90","volume":"14","author":"H Liu","year":"2024","unstructured":"Liu, H., Liu, J., Zhang, F., & Lu, W. (2024). Multi-level functional structure recognition of scientific literature. Journal of Information Resources Management, 14(3), 90\u2013103.","journal-title":"Journal of Information Resources Management"},{"key":"5355_CR38","doi-asserted-by":"publisher","first-page":"463","DOI":"10.1007\/s11192-018-2640-y","volume":"115","author":"W Lu","year":"2018","unstructured":"Lu, W., Huang, Y., Bu, Y., & Cheng, Q. (2018). Functional structure identification of scientific documents in computer science. Scientometrics, 115, 463\u2013486.","journal-title":"Scientometrics"},{"key":"5355_CR39","doi-asserted-by":"publisher","DOI":"10.1007\/s11192-025-05286-2","author":"W Lu","year":"2014","unstructured":"Lu, W., Huang, Y., & Cheng, Q. (2014). The structure function of academic text. Journal of the China Society for Scientific and Technical Information. https:\/\/doi.org\/10.1007\/s11192-025-05286-2","journal-title":"Journal of the China Society for Scientific and Technical Information"},{"key":"5355_CR40","doi-asserted-by":"crossref","unstructured":"Luo, F., Li, P., Zhou, J., Yang, P., Chang, B., Sui, Z., & Sun, X. (2019). A dual reinforcement learning framework for unsupervised text style transfer. arXiv preprint arXiv:1905.10060.","DOI":"10.24963\/ijcai.2019\/711"},{"issue":"6","key":"5355_CR41","first-page":"44","volume":"8","author":"J Mao","year":"2024","unstructured":"Mao, J., & Chen, Z. (2024). Identifying structural function of scientific literature abstracts based on deep active learning. Data Analysis and Knowledge Discovery, 8(6), 44\u201355.","journal-title":"Data Analysis and Knowledge Discovery"},{"issue":"2","key":"5355_CR42","doi-asserted-by":"publisher","first-page":"885","DOI":"10.1007\/s11192-021-04225-1","volume":"127","author":"B Ma","year":"2022","unstructured":"Ma, B., Zhang, C., Wang, Y., & Deng, S. (2022). Enhancing identification of structure function of academic articles using contextual information. Scientometrics, 127(2), 885\u2013925.","journal-title":"Scientometrics"},{"key":"5355_CR43","unstructured":"Nguyen, X.-P., Joty, S., Nguyen, T.-T., Wu, K., & Aw, A.T. (2021). Cross-model back-translated distillation for unsupervised machine translation. In: International Conference on Machine Learning, pp. 8073\u20138083 PMLR."},{"key":"5355_CR44","unstructured":"Peng, B., Li, C., He, P., Galley, M., & Gao, J. (2023). Instruction tuning with gpt-4. arXiv preprint arXiv:2304.03277."},{"issue":"11","key":"5355_CR45","first-page":"26","volume":"4","author":"C Qin","year":"2020","unstructured":"Qin, C., & Zhang, C. (2020). Recognizing structure functions of academic articles with hierarchical attention network. Data Analysis and Knowledge Discovery, 4(11), 26\u201342.","journal-title":"Data Analysis and Knowledge Discovery"},{"key":"5355_CR46","doi-asserted-by":"crossref","unstructured":"Sharma, S., Joshi, A., Zhao, Y., Mukhija, N., Bhathena, H., Singh, P., & Santhanam, S. (2023). When and how to paraphrase for named entity recognition? In: Proceedings of the 61st Annual Meeting of the Association for Computational Linguistics (Volume 1: Long Papers), pp. 7052\u20137087.","DOI":"10.18653\/v1\/2023.acl-long.390"},{"issue":"1","key":"5355_CR47","doi-asserted-by":"publisher","first-page":"101","DOI":"10.1186\/s40537-021-00492-0","volume":"8","author":"C Shorten","year":"2021","unstructured":"Shorten, C., Khoshgoftaar, T. M., & Furht, B. (2021). Text data augmentation for deep learning. Journal of Big Data, 8(1), 101.","journal-title":"Journal of Big Data"},{"issue":"3","key":"5355_CR48","first-page":"364","volume":"92","author":"LB Sollaci","year":"2004","unstructured":"Sollaci, L. B., & Pereira, M. G. (2004). The introduction, methods, results, and discussion (imrad) structure: A fifty-year survey. Journal of the Medical Library Association, 92(3), 364.","journal-title":"Journal of the Medical Library Association"},{"key":"5355_CR49","doi-asserted-by":"crossref","unstructured":"Wan, D., Zhang, Z., Zhu, Q., Liao, L., & Huang, M. (2022). A unified dialogue user simulator for few-shot data augmentation. In: Findings of the Association for Computational Linguistics: EMNLP 2022, pp. 3788\u20133799.","DOI":"10.18653\/v1\/2022.findings-emnlp.277"},{"key":"5355_CR50","doi-asserted-by":"crossref","unstructured":"Wang, X., Ji, H., Shi, C., Wang, B., Ye, Y., Cui, P., & Yu, P.S. (2019). Heterogeneous graph attention network. In: The World Wide Web Conference, 2022\u20132032 .","DOI":"10.1145\/3308558.3313562"},{"issue":"13","key":"5355_CR51","first-page":"95","volume":"63","author":"J Wang","year":"2019","unstructured":"Wang, J., Lu, W., Liu, J., & Cheng, Q. (2019). Research on structure function recognition of academic text based on multi-level fusion. Library and Information Service, 63(13), 95\u2013104.","journal-title":"Library and Information Service"},{"issue":"4","key":"5355_CR52","doi-asserted-by":"publisher","first-page":"1","DOI":"10.1145\/3464689","volume":"30","author":"H Wang","year":"2021","unstructured":"Wang, H., Xia, X., Lo, D., He, Q., Wang, X., & Grundy, J. (2021). Context-aware retrieval-based deep commit message generation. ACM Transactions on Software Engineering and Methodology (TOSEM), 30(4), 1\u201330.","journal-title":"ACM Transactions on Software Engineering and Methodology (TOSEM)"},{"key":"5355_CR53","doi-asserted-by":"crossref","unstructured":"Wei, J., & Zou, K. (2019). Eda: Easy data augmentation techniques for boosting performance on text classification tasks. arXiv preprint arXiv:1901.11196.","DOI":"10.18653\/v1\/D19-1670"},{"key":"5355_CR54","doi-asserted-by":"crossref","unstructured":"Wu, C., Ren, X., Luo, F., & Sun, X.: (2019). A hierarchical reinforced sequence operation method for unsupervised text style transfer. arXiv preprint arXiv:1906.01833.","DOI":"10.18653\/v1\/P19-1482"},{"key":"5355_CR55","unstructured":"Wu, F., Souza, A., Zhang, T., Fifty, C., Yu, T., & Weinberger, K. (2019). Simplifying graph convolutional networks. In: International Conference on Machine Learning, pp. 6861\u20136871 Pmlr."},{"issue":"1","key":"5355_CR56","doi-asserted-by":"publisher","first-page":"4","DOI":"10.1109\/TNNLS.2020.2978386","volume":"32","author":"Z Wu","year":"2020","unstructured":"Wu, Z., Pan, S., Chen, F., Long, G., Zhang, C., & Yu, P. S. (2020). A comprehensive survey on graph neural networks. IEEE Transactions on Neural Networks and Learning Systems, 32(1), 4\u201324.","journal-title":"IEEE Transactions on Neural Networks and Learning Systems"},{"key":"5355_CR57","unstructured":"Xu, W. (2023). China ranks first in the world in terms of the number of most influential journal papers in various disciplines. International Publishing Weekly. 001 edn, 25 September, 1\u20132."},{"issue":"4","key":"5355_CR58","doi-asserted-by":"publisher","DOI":"10.1016\/j.xinn.2021.100179","volume":"2","author":"Y Xu","year":"2021","unstructured":"Xu, Y., Liu, X., Cao, X., Huang, C., Liu, E., Qian, S., Liu, X., Wu, Y., Dong, F., Qiu, C.-W., et al. (2021). Artificial intelligence: A powerful paradigm for scientific research. The Innovation, 2(4), Article 100179.","journal-title":"The Innovation"},{"key":"5355_CR59","doi-asserted-by":"publisher","DOI":"10.1016\/j.eswa.2023.121909","volume":"238","author":"J Xu","year":"2024","unstructured":"Xu, J., Zhanyi, C. S., Xu, L., & Chen, L. (2024). Blendcse: Blend contrastive learnings for sentence embeddings with rich semantics and transferability. Expert Systems with Applications, 238, Article 121909.","journal-title":"Expert Systems with Applications"},{"key":"5355_CR60","doi-asserted-by":"crossref","unstructured":"Yang, Z., Yang, D., Dyer, C., He, X., Smola, A., & Hovy, E. (2016). Hierarchical attention networks for document classification. In: Proceedings of the 2016 Conference of the North American Chapter of the Association for Computational Linguistics: Human Language Technologies, 1480\u20131489.","DOI":"10.18653\/v1\/N16-1174"},{"key":"5355_CR61","unstructured":"Yao, X., Huang, Z., Hu, X., Yang, J., & Guo, Y. (2022). Masking the unknown: Leveraging masked samples for enhanced data augmentation. In: The 40th Conference on Uncertainty in Artificial Intelligence."},{"key":"5355_CR62","doi-asserted-by":"publisher","first-page":"7370","DOI":"10.1609\/aaai.v33i01.33017370","volume":"33","author":"L Yao","year":"2019","unstructured":"Yao, L., Mao, C., & Luo, Y. (2019). Graph convolutional networks for text classification. Proceedings of the AAAI Conference on Artificial Intelligence, 33, 7370\u20137377.","journal-title":"Proceedings of the AAAI Conference on Artificial Intelligence"},{"key":"5355_CR63","doi-asserted-by":"publisher","first-page":"259","DOI":"10.1162\/tacl_a_00097","volume":"4","author":"W Yin","year":"2016","unstructured":"Yin, W., Sch\u00fctze, H., Xiang, B., & Zhou, B. (2016). Abcnn: Attention-based convolutional neural network for modeling sentence pairs. Transactions of the Association for computational linguistics, 4, 259\u2013272.","journal-title":"Transactions of the Association for computational linguistics"},{"key":"5355_CR64","doi-asserted-by":"crossref","unstructured":"Zhang, J., Bao, K., Zhang, Y., Wang, W., Feng, F., & He, X. (2023). Is chatgpt fair for recommendation? evaluating fairness in large language model recommendation. In: Proceedings of the 17th ACM Conference on Recommender Systems, pp. 993\u2013999.","DOI":"10.1145\/3604915.3608860"},{"key":"5355_CR65","doi-asserted-by":"crossref","unstructured":"Zhang, B., Wang, N., Shao, Y., & Niu, Z. (2024). Hierarchy-aware bert-gcn dual-channel global model for hierarchical text classification. In: 2024 4th International Conference on Neural Networks, Information and Communication (NNICE), pp. 820\u2013825 IEEE.","DOI":"10.1109\/NNICE61279.2024.10498363"},{"key":"5355_CR66","first-page":"27429","volume":"10","author":"X Zhang","year":"2022","unstructured":"Zhang, X., Wu, L., Sun, J., Zhang, Q., & Lin, H. (2022). Comparative analysis of graph neural networks for node classification. IEEE Access, 10, 27429\u201327441.","journal-title":"IEEE Access"},{"key":"5355_CR67","doi-asserted-by":"publisher","DOI":"10.1016\/j.energy.2023.129654","volume":"287","author":"K Zhang","year":"2024","unstructured":"Zhang, K., Yang, X., Xu, L., Th\u00e9, J., Tan, Z., & Yu, H. (2024). Enhancing coal-gangue object detection using gan-based data augmentation strategy with dual attention mechanism. Energy, 287, Article 129654.","journal-title":"Energy"},{"issue":"6","key":"5355_CR68","first-page":"712","volume":"43","author":"Y Zhang","year":"2024","unstructured":"Zhang, Y., & Zhang, C. (2024). Identification of problem and method in scientific papers based on formulaic expression desensitization and enhanced boundary recognition. Journal of the China Society for Scientific and Technical Information, 43(6), 712\u2013732.","journal-title":"Journal of the China Society for Scientific and Technical Information"},{"issue":"9","key":"5355_CR69","first-page":"12","volume":"7","author":"Y Zhang","year":"2023","unstructured":"Zhang, Y., Zhang, C., Zhou, Y., & Chen, B. (2023). Chatgpt-based scientific paper entity recognition: Performance measurement and availability research. Data Analysis and Knowledge Discovery, 7(9), 12\u201324.","journal-title":"Data Analysis and Knowledge Discovery"},{"key":"5355_CR70","doi-asserted-by":"publisher","DOI":"10.48550\/arXiv.2502.05151","author":"H Zhang","year":"2023","unstructured":"Zhang, H., Zhao, Y., & Zhang, C. (2023). Recognition of research workflow paragraphs based on scibert and chatgpt data augmentation. Information Studies: Theory Application. https:\/\/doi.org\/10.48550\/arXiv.2502.05151","journal-title":"Information Studies: Theory Application"}],"container-title":["Scientometrics"],"original-title":[],"language":"en","link":[{"URL":"https:\/\/link.springer.com\/content\/pdf\/10.1007\/s11192-025-05355-6.pdf","content-type":"application\/pdf","content-version":"vor","intended-application":"text-mining"},{"URL":"https:\/\/link.springer.com\/article\/10.1007\/s11192-025-05355-6\/fulltext.html","content-type":"text\/html","content-version":"vor","intended-application":"text-mining"},{"URL":"https:\/\/link.springer.com\/content\/pdf\/10.1007\/s11192-025-05355-6.pdf","content-type":"application\/pdf","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2025,11,24]],"date-time":"2025-11-24T10:31:44Z","timestamp":1763980304000},"score":1,"resource":{"primary":{"URL":"https:\/\/link.springer.com\/10.1007\/s11192-025-05355-6"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2025,7,24]]},"references-count":70,"journal-issue":{"issue":"10","published-print":{"date-parts":[[2025,10]]}},"alternative-id":["5355"],"URL":"https:\/\/doi.org\/10.1007\/s11192-025-05355-6","relation":{},"ISSN":["0138-9130","1588-2861"],"issn-type":[{"type":"print","value":"0138-9130"},{"type":"electronic","value":"1588-2861"}],"subject":[],"published":{"date-parts":[[2025,7,24]]},"assertion":[{"value":"16 January 2025","order":1,"name":"received","label":"Received","group":{"name":"ArticleHistory","label":"Article History"}},{"value":"26 May 2025","order":2,"name":"accepted","label":"Accepted","group":{"name":"ArticleHistory","label":"Article History"}},{"value":"24 July 2025","order":3,"name":"first_online","label":"First Online","group":{"name":"ArticleHistory","label":"Article History"}}]}}