{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2026,7,12]],"date-time":"2026-07-12T16:26:14Z","timestamp":1783873574431,"version":"3.55.0"},"reference-count":60,"publisher":"Springer Science and Business Media LLC","issue":"3","license":[{"start":{"date-parts":[[2025,3,27]],"date-time":"2025-03-27T00:00:00Z","timestamp":1743033600000},"content-version":"tdm","delay-in-days":0,"URL":"https:\/\/creativecommons.org\/licenses\/by\/4.0"},{"start":{"date-parts":[[2025,3,27]],"date-time":"2025-03-27T00:00:00Z","timestamp":1743033600000},"content-version":"vor","delay-in-days":0,"URL":"https:\/\/creativecommons.org\/licenses\/by\/4.0"}],"funder":[{"DOI":"10.13039\/501100006692","name":"Universit\u00e0 degli Studi di Torino","doi-asserted-by":"crossref","id":[{"id":"10.13039\/501100006692","id-type":"DOI","asserted-by":"crossref"}]}],"content-domain":{"domain":["link.springer.com"],"crossmark-restriction":false},"short-container-title":["Lang Resources &amp; Evaluation"],"published-print":{"date-parts":[[2025,9]]},"abstract":"<jats:title>Abstract<\/jats:title>\n          <jats:p>Discourse representation structure (DRS), a formal meaning representation, has been used for both semantic parsing and natural language generation tasks and gained promising results for high-resource languages such as English and for the lesser-resourced European languages Italian, German, and Dutch. We investigate how we can employ DRS for the low-resource language Urdu for neural semantic parsing (translating Urdu sentences into formal meaning representations) and natural language generation (generating Urdu sentences from formal meaning representations). There are no annotated corpora for Urdu available, so we adopted a combined approach involving both manual annotations and rule-based procedures to transform English-aligned DRS into Urdu-aligned DRS through syntactic structure and word surface alignment, because word order in Urdu (subject\u2013object\u2013verb) differs from that of English (subject\u2013verb\u2013object). To further increase the amount of semantically annotated data, we developed lexical, grammatical, and named entity-based augmentation techniques. This resulted in an increase of nine times more data examples. Using the augmented meaning bank for Urdu, we developed a neural semantic parser and generator that benefited significantly from the augmented data and showed more generalization ability compared to the model without augmentation. We evaluated the effect of semantic data augmentation using a transformer-based state-of-the-art neural sequence-to-sequence architecture. Our implementation shows promising results for the semantic processing of Urdu and demonstrates that data augmentation increases performance (F1-Score) for semantic parsing from 67.12 to 76.81, and leads to substantially increased BLEU, BERT-Score, METEOR, ROUGE, and chrF scores for generation.<\/jats:p>","DOI":"10.1007\/s10579-025-09819-2","type":"journal-article","created":{"date-parts":[[2025,3,30]],"date-time":"2025-03-30T08:11:43Z","timestamp":1743322303000},"page":"2469-2500","update-policy":"https:\/\/doi.org\/10.1007\/springer_crossmark_policy","source":"Crossref","is-referenced-by-count":3,"title":["Semantic processing for Urdu: corpus creation, parsing, and generation"],"prefix":"10.1007","volume":"59","author":[{"given":"Muhammad Saad","family":"Amin","sequence":"first","affiliation":[],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Xiao","family":"Zhang","sequence":"additional","affiliation":[],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Luca","family":"Anselma","sequence":"additional","affiliation":[],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Alessandro","family":"Mazzei","sequence":"additional","affiliation":[],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Johan","family":"Bos","sequence":"additional","affiliation":[],"role":[{"vocabulary":"crossref","role":"author"}]}],"member":"297","published-online":{"date-parts":[[2025,3,27]]},"reference":[{"issue":"2","key":"9819_CR1","first-page":"45","volume":"25","author":"L Abzianidze","year":"2020","unstructured":"Abzianidze, L., van Noord, R., Wang, C., & Bos, J. (2020). The parallel meaning bank: A framework for semantically annotating multiple languages.\nApplied Mathematics and Informatics ,\n25(2), 45\u201360.","journal-title":"Applied Mathematics and Informatics"},{"key":"9819_CR2","unstructured":"Ahmed, T., & Hautli, A. (2010). Developing a basic lexical resource for Urdu using Hindi wordnet. In: Proceedings of CLT10, Islamabad, Pakistan."},{"key":"9819_CR3","unstructured":"Amin, M.S., Anselma, L., & Mazzei, A. (2024a). Data augmentation for low-resource italian nlp: Enhancing semantic processing with drs. In: Proceedings of the Tenth iIalian Conference on Computational Linguistics (clic-it 2024) (pp. 1\u201310). Pisa, Italy. Retrieved from https:\/\/ceur-ws.org\/Vol-3878\/5 main long.pdf"},{"key":"9819_CR4","doi-asserted-by":"crossref","unstructured":"Amin, M.S., Anselma, L., & Mazzei, A. (2024b). Exploring data augmentation in neural DRS-to-text generation. In: Y. Graham &M. Purver (Eds.), Proceedings of the 18th Conference of the European Chapter of the Association for Computational Linguistics (volume 1: Long papers) (pp. 2164\u20132178). St. Julian\u2019s, Malta: Association for Computational Linguistics. Retrieved from https:\/\/aclanthology.org\/2024.eacl-long.132","DOI":"10.18653\/v1\/2024.eacl-long.132"},{"key":"9819_CR5","unstructured":"Amin, M.S., Mazzei, A., & Anselma, L., et al. (2022). Towards data augmentation for drs-to-text generation. In: Ceur Workshop Proceedings (Vol. 3287, pp. 141\u2013152)."},{"key":"9819_CR6","volume-title":"Logics of conversation","author":"N Asher","year":"2003","unstructured":"Asher, N., & Lascarides, A. (2003). Logics of conversation. Cambridge University Press."},{"key":"9819_CR7","unstructured":"Banarescu, L., Bonial, C., Cai, S., Georgescu, M., Griffitt, K.,  Hermjakob, U., Knight, K., Kohen, P., Palmer, M., &. Schneider, N. (2013). Abstract meaning representation for sembanking. In: Proceedings of the 7th Linguistic Annotation Workshop and Interoperability with Discourse (pp. 178\u2013186)."},{"key":"9819_CR9","unstructured":"Banerjee, S., & Lavie, A. (2005). Meteor: An automatic metric for mt evaluation with improved correlation with human judgments. In: Proceedings of the ACL Workshop On Intrinsic And Extrinsic Evaluation Measures For Machine Translation And\/Or Summarization (pp. 65\u201372)."},{"key":"9819_CR10","unstructured":"Basile, V., & Bos, J. (2011a). Towards generating text from discourse representation structures. In: Enlg\u201911 proceedings of the 13th European Workshop on Natural Language Generation, (pp. 145\u2013150)."},{"key":"9819_CR11","unstructured":"Basile, V., Bos, J., Evang, K., & Venhuizen, N. (2012). Developing a large semantically annotated corpus. In: Lrec 2012, Eighth International Conference On Language Resources And Evaluation."},{"key":"9819_CR12","first-page":"145","volume":"11","author":"V Basile","year":"2011","unstructured":"Basile, V., & Bos, J. (2011). Towards generating text from discourse representation structures. ENLG\u2019, 11, 145\u2013150.","journal-title":"ENLG\u2019"},{"key":"9819_CR13","doi-asserted-by":"crossref","unstructured":"Belouadi, J., & Eger, S. (2023). ByGPT5: End-to-end style-conditioned poetry generation with token-free language models. A. Rogers, J. Boyd-Graber, & N. Okazaki (Eds.), Proceedings of the 61st Annual Meeting of the Association for Computational Linguistics (volume 1: Long papers) (pp. 7364\u20137381). Toronto, Canada: Association for Computational Linguistics. Retrieved from https:\/\/aclanthology.org\/2023.acl-long.406","DOI":"10.18653\/v1\/2023.acl-long.406"},{"key":"9819_CR15","unstructured":"B\u00f6gel, T., Butt, M., Hautli, A., & Sulger, S. (2008). Developing a finite-state morphological analyzer for urdu and hindi. Universit\u00e4t Potsdam."},{"key":"9819_CR16","unstructured":"B\u00f6gel, T., Butt, M., Hautli, A., & Sulger, S. (2009). Urdu and the modular architecture of pargram. In: Proceedings of the Conference on Language and Technology (Vol. 70)."},{"key":"9819_CR17","doi-asserted-by":"publisher","first-page":"155","DOI":"10.1146\/annurev-linguistics-062419-125014","volume":"6","author":"K B\u00f6rjars","year":"2020","unstructured":"B\u00f6rjars, K. (2020). Lexical-functional grammar: an overview. Annual Review of Linguistics, 6, 155\u2013172.","journal-title":"Annual Review of Linguistics"},{"key":"9819_CR18","doi-asserted-by":"crossref","unstructured":"Bos, J. (2008). Wide-coverage semantic analysis with Boxer. In: Semantics in text processing. STEP 2008 Conference Proceedings (pp. 277\u2013286). College Publications. Retrieved from https:\/\/aclanthology.org\/W08-2222","DOI":"10.3115\/1626481.1626503"},{"key":"9819_CR19","unstructured":"Bos, J. (2023). The sequence notation: Catching complex meanings in simple graphs. Proceedings of the 15th international conference on computational semantics (iwcs 2023) (pp. 1\u201314). Nancy, France."},{"key":"9819_CR20","doi-asserted-by":"crossref","unstructured":"Butt, M., & King, T.H. (2002). Urdu and the parallel grammar project. Coling-02: The 3rd workshop on asian language resources and international standardization","DOI":"10.3115\/1118759.1118762"},{"key":"9819_CR21","unstructured":"Cai, S., & Knight, K. (2013). Smatch: an evaluation metric for semantic feature structures. H. Schuetze, P. Fung, & M. Poesio (Eds.), Proceedings of the 51st Annual Meeting of the Association for Computational Linguistics (volume 2: Short papers) (pp. 748\u2013752). Association for Computational Linguistics. Retrieved from https:\/\/aclanthology.org\/P13-2131"},{"key":"9819_CR22","doi-asserted-by":"crossref","unstructured":"Copestake, A. (2009). Invited Talk: slacker semantics: Why superficiality, dependency and avoidance of commitment can be the right way to go. In: A. Lascarides, C. Gardent, & J. Nivre (Eds.), Proceedings of the 12th Cnference of the European Chapter of the ACL (EACL 2009) (pp. 1\u20139). Association for Computational Linguistics. Retrieved from https:\/\/aclanthology.org\/E09-1001\/","DOI":"10.3115\/1609067.1609167"},{"key":"9819_CR23","unstructured":"Enis, M., & Hopkins, M. (2024). From LLM to NMT: Advancing low-resource machine translation with Claude. ArXiv, abs\/2404.13813, , Retrieved from https:\/\/api.semanticscholar.org\/CorpusID:269292906"},{"key":"9819_CR24","doi-asserted-by":"crossref","unstructured":"Fancellu, F., Gilroy, S., Lopez, A., & Lapata, M. (2019). Semantic graph parsing with recurrent neural network DAG grammars. In: Proceedings of the 2019 Conference on Empirical Methods in Natural Language Processing and the 9th International Joint Conference on Natural Language Processing (emnlp-ijcnlp) (pp. 2769\u20132778). Hong Kong, China: Association for Computational Linguistics. Retrieved from https:\/\/aclanthology.org\/D19-1278","DOI":"10.18653\/v1\/D19-1278"},{"key":"9819_CR25","doi-asserted-by":"crossref","unstructured":"Feng, S.Y., Gangal, V., Wei, J., Chandar, S., Vosoughi, S., Mitamura, T., & Hovy, E. (2021). A survey of data augmentation approaches for NLP. Findings of the association for computational linguistics: Acl-ijcnlp 2021 (pp. 968\u2013988). Online: Association for Computational Linguistics. Retrieved from https:\/\/aclanthology.org\/2021.findingsacl 84","DOI":"10.18653\/v1\/2021.findings-acl.84"},{"key":"9819_CR26","doi-asserted-by":"publisher","unstructured":"Flanigan, J., Dyer, C., Smith, N.A., & Carbonell, J.G. (2016). Generation from abstract meaning representation using tree transducers. in Proc., 2016, 731\u2013739, https:\/\/doi.org\/10.18653\/v1\/N16-1087]","DOI":"10.18653\/v1\/N16-1087"},{"key":"9819_CR27","unstructured":"Frege, G. (1884). Die grundlagen der arithmetik: Eine logisch-mathematische untersuchung \u00fcber den begriff der zahl. Breslau: W. Koebner. (Translated as *The Foundations of Arithmetic: A Logico-Mathematical Enquiry into the Concept of Number*, 1950, by J. L. Austin, Blackwell.)"},{"key":"9819_CR28","doi-asserted-by":"crossref","unstructured":"Fu, Q., Zhang, Y., Liu, J., & Zhang, M. (2020). DRTS parsing with structure-aware encoding and decoding. In: Proceedings of the 58th Annual Meeting of the Association for Computational Linguistics (pp. 6818\u20136828). Online: Association for Computational Linguistics. Retrieved from https:\/\/aclanthology.org\/2020.acl-main.609","DOI":"10.18653\/v1\/2020.acl-main.609"},{"key":"9819_CR29","unstructured":"Hautli, A., & Butt, M. (2011). Towards a computational semantic analyzer for urdu. In: Proceedings of the 9th Workshop on Asian Language Resources (pp. 71\u201378)."},{"issue":"8","key":"9819_CR30","doi-asserted-by":"publisher","first-page":"1735","DOI":"10.1162\/neco.1997.9.8.1735","volume":"9","author":"S Hochreiter","year":"1997","unstructured":"Hochreiter, S., & Schmidhuber, J. (1997). Long short-term memory. Neural Computation, 9(8), 1735\u20131780.","journal-title":"Neural Computation"},{"key":"9819_CR31","doi-asserted-by":"crossref","unstructured":"In\u00b4acio, M., & Pardo, T. (2021). Semantic-based opinion summarization. In: R. Mitkov & G. Angelova (Eds.), Proceedings of the International Conference on Recent Advances in Natural language Processing (ranlp 2021) (pp. 619\u2013628). Held Online: INCOMA Ltd. Retrieved from https:\/\/aclanthology.org\/2021.ranlp-1.70","DOI":"10.26615\/978-954-452-072-4_070"},{"key":"9819_CR32","doi-asserted-by":"publisher","unstructured":"Jafar, M.R., & Jafar, M.R. (2022). The challenges toward the implications of official Urdu: Challenges toward the implication of official Urdu language. Pacific International Journal, 5(4), 33\u201337.  https:\/\/doi.org\/10.55014\/pij.v5i4.234 Retrieved from https:\/\/rclss.com\/pij\/article\/view\/234","DOI":"10.55014\/pij.v5i4.234"},{"key":"9819_CR33","doi-asserted-by":"crossref","unstructured":"Kasper, R.T. (1989). A flexible interface for linking applications to Penman\u2019s sentence generator. In: Speech and Natural Language: Proceedings of a Workshop Held at Philadelphia, Pennsylvania, February 21\u201323, 1989. Retrieved from https:\/\/aclanthology.org\/H89- 1022","DOI":"10.3115\/100964.100979"},{"key":"9819_CR34","unstructured":"Kleinecke, D. (1986). The mental representation of grammatical relations. Computational Linguistics, 12(2),"},{"key":"9819_CR35","doi-asserted-by":"crossref","unstructured":"Lewis, M., Liu, Y., Goyal, N., Ghazvininejad, M., Mohamed, A., Levy, O., Stoyanov, V., & Zettlemoyer, L. (2020). BART: Denoising sequence-to-sequence pre-training for natural language generation, translation, and comprehension. In: Proceedings of the 58th Annual Meeting of the Association for Computational Linguistics (pp. 7871\u20137880). Online: Association for Computational Linguistics. Retrieved from https:\/\/aclanthology.org\/2020.aclmain.703.","DOI":"10.18653\/v1\/2020.acl-main.703"},{"key":"9819_CR36","doi-asserted-by":"crossref","unstructured":"Li, B., Wen, Y., Qu, W., Bu, L., & Xue, N. (2016). Annotating the little prince with chinese amrs. In: Proceedings of the 10th Linguistic Annotation Workshop Held in Conjunction with ACL 2016 (law-x 2016) (pp. 7\u201315).","DOI":"10.18653\/v1\/W16-1702"},{"key":"9819_CR37","unstructured":"Lin, C.-Y. (2004). ROUGE: A package for automatic evaluation of summaries. Text summarization branches out (pp. 74\u201381). Association for Computational Linguistics. Retrieved from https:\/\/aclanthology.org\/W04-1013"},{"key":"9819_CR38","doi-asserted-by":"crossref","unstructured":"Liu, J., Cohen, S.B., & Lapata, M. (2018). Discourse representation structure parsing. In: Proceedings of the 56th Annual Meeting of the Association for Computational Linguistics (volume 1: Long papers) (pp. 429\u2013439). Association for Computational Linguistics. Retrieved from https:\/\/aclanthology.org\/P18-1040","DOI":"10.18653\/v1\/P18-1040"},{"key":"9819_CR39","doi-asserted-by":"crossref","unstructured":"Liu, J., Cohen, S.B., & Lapata, M. (2019). Discourse representation parsing for sentences and documents. In: Proceedings of the 57th Annual Meeting of the Association for Computational Linguistics (pp. 6248\u20136262). Association for Computational Linguistics. Retrieved from https:\/\/aclanthology.org\/P19-1629","DOI":"10.18653\/v1\/P19-1629"},{"key":"9819_CR40","doi-asserted-by":"crossref","unstructured":"Liu, J., Cohen, S.B., & Lapata, M. (2021). Text generation from discourse representation structures. In: Proceedings of the 2021 Conference of the North American Chapter of the Association for Computational Linguistics: Human Language Technologies (pp. 397\u2013415).","DOI":"10.18653\/v1\/2021.naacl-main.35"},{"key":"9819_CR41","unstructured":"Noord, R.v. (2019). Neural boxer at the iwcs shared task on drs parsing. Gothenburg, Sweden. Association for Computational Linguistics: in Proc. IWCS Shared Task on Semantic Parsing."},{"key":"9819_CR42","doi-asserted-by":"crossref","unstructured":"Papineni, K., Roukos, S., Ward, T., & Zhu, W.-J. (2002). Bleu: a method for automatic evaluation of machine translation. In: Proceedings of the 40th Annual Meeting of the Association for Computational Linguistics (pp. 311\u2013318).","DOI":"10.3115\/1073083.1073135"},{"key":"9819_CR43","doi-asserted-by":"publisher","unstructured":"Pasha, A.R., Abbas, N., & Ahmed, H.N. (2022). Epenthesis: The movement of the urdu alveolar-fricative sound into the punjabi palatal-affricate sound. Engineering, Technology & Applied Science Research, 12(6), 9487\u20139490, https:\/\/doi.org\/10.48084\/etasr.5295.","DOI":"10.48084\/etasr.5295"},{"key":"9819_CR44","unstructured":"Poelman, W., van Noord, R., & Bos, J. (2022, October). Transparent semantic parsing with Universal Dependencies using graph transformations. Proceedings of the 29th international conference on computational linguistics (pp. 4186\u20134192). Gyeongju, Republic of Korea: International Committee on Computational Linguistics. Retrieved from https:\/\/aclanthology.org\/2022.coling-1.367"},{"key":"9819_CR45","doi-asserted-by":"crossref","unstructured":"Popovi\u0107, M. (2015). chrF: character n-gram F-score for automatic MT evaluation. O. Bojar et al. (Eds.), Proceedings of the Tenth Workshop on Statistical Machine Translation (pp. 392\u2013395). Association for Computational Linguistics. Retrieved from https:\/\/aclanthology.org\/W15-3049","DOI":"10.18653\/v1\/W15-3049"},{"key":"9819_CR46","unstructured":"Raffel, C., Shazeer, N., Roberts, A., Lee, K., Narang, S., Matena, M., Zhou, Y., Li, W., &  Liu, P.J. (2020). Exploring the limits of transfer learning with a unified text-to-text transformer. Journal of Machine Learning Research, 21(1)."},{"key":"9819_CR47","unstructured":"Schuler, K.K., & Palmer, M. (2005). Verbnet: a broad-coverage, comprehensive verb lexicon.. Retrieved from https:\/\/api.semanticscholar.org\/CorpusID:60771008"},{"issue":"1","key":"9819_CR48","doi-asserted-by":"publisher","first-page":"1","DOI":"10.1186\/s40537-019-0197-0","volume":"6","author":"C Shorten","year":"2019","unstructured":"Shorten, C., & Khoshgoftaar, T. M. (2019). A survey on image data augmentation for deep learning. Journal of Big Data, 6(1), 1\u201348.","journal-title":"Journal of Big Data"},{"key":"9819_CR49","doi-asserted-by":"crossref","unstructured":"Shou, Z., Jiang, Y., & Lin, F. (2022). AMR-DA: Data augmentation by Abstract Meaning Representation. In: S. Muresan, P. Nakov, & A. Villavicencio (Eds.), Findings of the association for computational linguistics: Acl 2022 (pp. 3082\u2013 3098). Association for Computational Linguistics. Retrieved from https:\/\/aclanthology.org\/2022.findings-acl.244","DOI":"10.18653\/v1\/2022.findings-acl.244"},{"issue":"5","key":"9819_CR50","doi-asserted-by":"publisher","first-page":"2636","DOI":"10.3390\/app12052636","volume":"12","author":"L Stankevi\u010dius","year":"2022","unstructured":"Stankevi\u010dius, L., Luko\u0161evi\u010dius, M., Kapo\u010di\u016bt\u0117-Dzikien\u0117, J., Briedien\u0117, M., & Krilavi\u010dius, T. (2022). Correcting diacritics and typos with a byt5 transformer model. Applied Sciences, 12(5), 2636.","journal-title":"Applied Sciences"},{"key":"9819_CR51","unstructured":"Touvron, H., Lavril, T., Izacard, G., Martinet, X., Lachaux, M.-A., Lacroix, T., Rozi\u00e8re, B., Goyal, N., Hambro, E., Azhar, F., Rodriguez, A., Joulin, A., Grave, E., & Lample, G. (2023). Llama: Open and efficient foundation language models. ArXiv, abs\/2302.13971, , Retrieved from https:\/\/api.semanticscholar.org\/CorpusID:257219404"},{"key":"9819_CR52","unstructured":"van Noord, R., Abzianidze, L., Haagsma, H., & Bos, J. (2018). Evaluating scoped meaning representations. In: Proceedings of the Eleventh International Conference on Language Resources and Evaluation (LREC 2018). European Language Resources Association (ELRA). Retrieved from https:\/\/aclanthology.org\/L18-1267"},{"key":"9819_CR53","doi-asserted-by":"crossref","unstructured":"van Noord, R., Toral, A., & Bos, J. (2020). Character-level representations improve DRS-based semantic parsing even in the age of BERT. In: Proceedings of the 2020 Conference on Empirical Methods in Natural Language Processing (EMNLP) (pp. 4587\u20134603). Association for Computational Linguistics. Retrieved from https:\/\/aclanthology.org\/2020.emnlp-main.371","DOI":"10.18653\/v1\/2020.emnlp-main.371"},{"key":"9819_CR54","doi-asserted-by":"crossref","unstructured":"Wang, C., Lai, H., Nissim, M., & Bos, J. (2023). Pre-trained language-meaning models for multilingual parsing and generation. Findings of the association for computational linguistics: Acl 2023 (pp. 5586\u20135600). Association for Computational Linguistics. Retrieved from https:\/\/aclanthology.org\/2023.findings-acl.345","DOI":"10.18653\/v1\/2023.findings-acl.345"},{"key":"9819_CR55","doi-asserted-by":"crossref","unstructured":"Wang, C., van Noord, R., Bisazza, A., & Bos, J. (2021a). Evaluating text generation from discourse representation structures. In: Proceedings of the 1st Workshop on Natural Language Generation, Evaluation, and Metrics (gem 2021) (pp. 73\u201383). Online: Association for Computational Linguistics. Retrieved from https:\/\/aclanthology.org\/2021.gem-1.8","DOI":"10.18653\/v1\/2021.gem-1.8"},{"key":"9819_CR57","doi-asserted-by":"crossref","unstructured":"Wang, C., van Noord, R., Bisazza, A., & Bos, J. (2021b). Input representations for parsing discourse representation structures: Comparing English with Chinese. C. Zong, F. Xia, W. Li, & R. Navigli (Eds.), Proceedings of the 59th Annual Meeting of the Association for Computational Linguistics and the 11th International Joint Conference on Natural Language Processing (volume 2: Short papers) (pp. 767\u2013775). Online: Association for Computational Linguistics. Retrieved from https:\/\/aclanthology.org\/2021.aclshort 97","DOI":"10.18653\/v1\/2021.acl-short.97"},{"key":"9819_CR58","doi-asserted-by":"crossref","unstructured":"Xue, L., Barua, A., Constant, N., Al-Rfou, R., Narang, S., & Kale, M., Roberts, A., & Raffel, C. (2022). Byt5: Towards a token-free future with pre-trained byte-to-byte models. Transactions of the Association for Computational Linguistics,10, 291\u2013306.","DOI":"10.1162\/tacl_a_00461"},{"key":"9819_CR59","doi-asserted-by":"crossref","unstructured":"Xue, L., Constant, N., Roberts, A., Kale, M., Al-Rfou, R., Siddhant, A., Barua, A., & Raffel, C. (2021). mT5: A massively multilingual pre-trained text-to-text transformer. K. Toutanova et al. (Eds.), Proceedings of the 2021 Conference of the North American Chapter of the Association for Computational Linguistics: Human Language Technologies (pp. 483\u2013498). Online: Association for Computational Linguistics. Retrieved from https:\/\/aclanthology.org\/2021.naacl-main.41","DOI":"10.18653\/v1\/2021.naacl-main.41"},{"key":"9819_CR60","doi-asserted-by":"crossref","unstructured":"Yih, W.-t., He, X., & Meek, C. (2014). Semantic parsing for single-relation question answering. In: Proceedings of the 52nd Annual Meeting of the Association for Computational Linguistics (volume 2: Short papers) (pp. 643\u2013648).","DOI":"10.3115\/v1\/P14-2105"},{"key":"9819_CR61","unstructured":"Yin, Y., Li, Y., Meng, F., Zhou, J., & Zhang, Y. (2022). Categorizing semantic representations for neural machine translation. N. Calzolari et al. (Eds.), Proceedings of the 29th International Conference on Computational Linguistics (pp. 5227\u20135239). International Committee on Computational Linguistics. Retrieved from https:\/\/aclanthology.org\/2022.coling-1.464"},{"key":"9819_CR62","unstructured":"Zhang, T., Kishore, V., Wu, F., Weinberger, K.Q., & Artzi, Y. (2020). Bertscore: Evaluating text generation with BERT. In: 8th International Conference on Learning Representations, ICLR 2020, Addis ababa, Ethiopia, April 26-30, 2020. OpenReview.net. Retrieved from https:\/\/openreview.net\/forum?id=SkeHuCVFDr"},{"key":"9819_CR63","doi-asserted-by":"crossref","unstructured":"Zhong, W., Xu, J., Tang, D., Xu, Z., Duan, N., Zhou, M., Wang, J., & Yin, J. (2020). Reasoning over semantic-level graph for fact checking. In: D. Jurafsky, J. Chai, N. Schluter, & J. Tetreault (Eds.), Proceedings of the 58th Annual Meeting of the Association for Computational Linguistics (pp. 6170\u20136180). Online: Association for Computational Linguistics. Retrieved from https:\/\/aclanthology.org\/2020.acl-main.549","DOI":"10.18653\/v1\/2020.acl-main.549"}],"container-title":["Language Resources and Evaluation"],"original-title":[],"language":"en","link":[{"URL":"https:\/\/link.springer.com\/content\/pdf\/10.1007\/s10579-025-09819-2.pdf","content-type":"application\/pdf","content-version":"vor","intended-application":"text-mining"},{"URL":"https:\/\/link.springer.com\/article\/10.1007\/s10579-025-09819-2\/fulltext.html","content-type":"text\/html","content-version":"vor","intended-application":"text-mining"},{"URL":"https:\/\/link.springer.com\/content\/pdf\/10.1007\/s10579-025-09819-2.pdf","content-type":"application\/pdf","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2025,9,6]],"date-time":"2025-09-06T08:22:49Z","timestamp":1757146969000},"score":1,"resource":{"primary":{"URL":"https:\/\/link.springer.com\/10.1007\/s10579-025-09819-2"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2025,3,27]]},"references-count":60,"journal-issue":{"issue":"3","published-print":{"date-parts":[[2025,9]]}},"alternative-id":["9819"],"URL":"https:\/\/doi.org\/10.1007\/s10579-025-09819-2","relation":{},"ISSN":["1574-020X","1574-0218"],"issn-type":[{"value":"1574-020X","type":"print"},{"value":"1574-0218","type":"electronic"}],"subject":[],"published":{"date-parts":[[2025,3,27]]},"assertion":[{"value":"4 March 2025","order":1,"name":"accepted","label":"Accepted","group":{"name":"ArticleHistory","label":"Article History"}},{"value":"27 March 2025","order":2,"name":"first_online","label":"First Online","group":{"name":"ArticleHistory","label":"Article History"}},{"order":1,"name":"Ethics","group":{"name":"EthicsHeading","label":"Declarations"}},{"value":"The authors have no Conflict of interest to declare that are relevant to the content of this article.","order":2,"name":"Ethics","group":{"name":"EthicsHeading","label":"Conflict of interest"}}]}}