{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2026,5,7]],"date-time":"2026-05-07T15:52:02Z","timestamp":1778169122419,"version":"3.51.4"},"reference-count":35,"publisher":"Association for Computing Machinery (ACM)","issue":"8","license":[{"start":{"date-parts":[[2024,8,8]],"date-time":"2024-08-08T00:00:00Z","timestamp":1723075200000},"content-version":"vor","delay-in-days":0,"URL":"https:\/\/www.acm.org\/publications\/policies\/copyright_policy#Background"}],"content-domain":{"domain":["dl.acm.org"],"crossmark-restriction":true},"short-container-title":["ACM Trans. Asian Low-Resour. Lang. Inf. Process."],"published-print":{"date-parts":[[2024,8,31]]},"abstract":"<jats:p>\n            Common spelling checks in the current digital era have trouble reading languages such as Bengali, which employ English letters differently. In response, we have created a better Bidirectional Encoder Representations from Transformers (BERT)\u2013based spell checker that makes use of a convolutional neural network (CNN) sub-model (Semantic Network). Our novelty, which we term\n            <jats:italic>progressive stacking<\/jats:italic>\n            , concentrates on improving BERT model training while expediting the corrective process. We discovered that, when comparing shallow and deep versions, deeper models could require less training time. There is potential for improving spelling corrections with this technique. We categorized and utilized as a test set a 6,300-word dataset that Nayadiganta Mohiuddin supplied, some of which had spelling errors. The most popular terms were the same as those found in the Prothom-Alo artificial error dataset.\n          <\/jats:p>","DOI":"10.1145\/3669941","type":"journal-article","created":{"date-parts":[[2024,7,5]],"date-time":"2024-07-05T11:13:52Z","timestamp":1720178032000},"page":"1-12","update-policy":"https:\/\/doi.org\/10.1145\/crossmark-policy","source":"Crossref","is-referenced-by-count":1,"title":["BERT-Inspired Progressive Stacking to Enhance Spelling Correction in Bengali Text"],"prefix":"10.1145","volume":"23","author":[{"ORCID":"https:\/\/orcid.org\/0000-0002-3756-864X","authenticated-orcid":false,"given":"Debajyoty","family":"Banik","sequence":"first","affiliation":[{"name":"School of Computer Science and Artificial Intelligence, SR University, Warangal, India"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"ORCID":"https:\/\/orcid.org\/0009-0000-2751-1490","authenticated-orcid":false,"given":"Saneyika","family":"Das","sequence":"additional","affiliation":[{"name":"Kalinga Institute of Industrial Technology Deemed to be University, Bhubaneswar, India"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"ORCID":"https:\/\/orcid.org\/0000-0003-1475-3962","authenticated-orcid":false,"given":"SHESHIKALA","family":"MARTHA","sequence":"additional","affiliation":[{"name":"SR University, Warangal, India"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"ORCID":"https:\/\/orcid.org\/0000-0003-3165-3293","authenticated-orcid":false,"given":"Achyut","family":"Shankar","sequence":"additional","affiliation":[{"name":"Department of Cyber Systems Engineering, WMG, University of Warwick, Coventry, United Kingdom, Center of Research Impact and Outcome, Chitkara University, Punjab, India, University Centre for Research &amp; Development, Chandigarh University, Mohali, India, and Department of Computer Science and Engineering, Graphic Era Deemed to be University, Dehradun, India"}],"role":[{"role":"author","vocabulary":"crossref"}]}],"member":"320","published-online":{"date-parts":[[2024,8,8]]},"reference":[{"key":"e_1_3_2_2_2","doi-asserted-by":"publisher","DOI":"10.1016\/j.cie.2022.108534"},{"key":"e_1_3_2_3_2","first-page":"103","volume-title":"Proceedings of COLING 2012: Posters","author":"Attia Mohammed","year":"2012","unstructured":"Mohammed Attia, Pavel Pecina, Younes Samih, Khaled Shaalan, and Josef Van Genabith. 2012. Improved spelling error detection and correction for Arabic. In Proceedings of COLING 2012: Posters. 103\u2013112."},{"key":"e_1_3_2_4_2","doi-asserted-by":"publisher","DOI":"10.1017\/S1351324915000030"},{"key":"e_1_3_2_5_2","volume-title":"Handling Arabic Morphological and Syntactic Ambiguity within the LFG Framework with a View to Machine Translation","author":"Attia Mohammed A.","year":"2008","unstructured":"Mohammed A. Attia. 2008. Handling Arabic Morphological and Syntactic Ambiguity within the LFG Framework with a View to Machine Translation. The University of Manchester (United Kingdom)."},{"key":"e_1_3_2_6_2","doi-asserted-by":"publisher","DOI":"10.1016\/j.est.2020.101466"},{"key":"e_1_3_2_7_2","doi-asserted-by":"publisher","DOI":"10.1109\/ISED.2016.7977080"},{"key":"e_1_3_2_8_2","doi-asserted-by":"publisher","DOI":"10.1007\/s12046-020-01427-w"},{"key":"e_1_3_2_9_2","doi-asserted-by":"publisher","DOI":"10.1016\/j.asoc.2019.02.031"},{"key":"e_1_3_2_10_2","doi-asserted-by":"publisher","DOI":"10.1016\/j.heliyon.2019.e02504"},{"key":"e_1_3_2_11_2","first-page":"10","volume-title":"Proceedings of the 13th International Conference on Natural Language Processing","author":"Banik Debajyoty","year":"2016","unstructured":"Debajyoty Banik, Sukanta Sen, Asif Ekbal, and Pushpak Bhattacharyya. 2016. Can SMT and RBMT improve each other\u2019s performance?\u2014An experiment with English-Hindi translation. In Proceedings of the 13th International Conference on Natural Language Processing. 10\u201319."},{"key":"e_1_3_2_12_2","first-page":"3550","volume-title":"Proceedings of the 9th International Conference on Language Resources and Evaluation (LREC\u201914)","author":"Bojar Ond\u0159ej","year":"2014","unstructured":"Ond\u0159ej Bojar, Vojt\u011bch Diatka, Pavel Rychl\u1ef3, Pavel Stra\u0148\u00e1k, V\u00edt Suchomel, Ale\u0161 Tamchyna, and Daniel Zeman. 2014. Hindencorp-Hindi-English and Hindi-only corpus for machine translation. In Proceedings of the 9th International Conference on Language Resources and Evaluation (LREC\u201914). 3550\u20133555."},{"key":"e_1_3_2_13_2","article-title":"SpellGCN: Incorporating phonological and visual similarities into language models for Chinese spelling check","author":"Cheng Xingyi","year":"2020","unstructured":"Xingyi Cheng, Weidi Xu, Kunlong Chen, Shaohua Jiang, Feng Wang, Taifeng Wang, Wei Chu, and Yuan Qi. 2020. SpellGCN: Incorporating phonological and visual similarities into language models for Chinese spelling check. arXiv preprint arXiv:2004.14166 (2020).","journal-title":"arXiv preprint arXiv:2004.14166"},{"key":"e_1_3_2_14_2","article-title":"Universal transformers","author":"Dehghani Mostafa","year":"2018","unstructured":"Mostafa Dehghani, Stephan Gouws, Oriol Vinyals, Jakob Uszkoreit, and \u0141ukasz Kaiser. 2018. Universal transformers. arXiv preprint arXiv:1807.03819 (2018).","journal-title":"arXiv preprint arXiv:1807.03819"},{"key":"e_1_3_2_15_2","doi-asserted-by":"publisher","DOI":"10.1109\/OCIT56763.2022.00031"},{"key":"e_1_3_2_16_2","article-title":"BERT: Pre-training of deep bidirectional transformers for language understanding","author":"Devlin Jacob","year":"2018","unstructured":"Jacob Devlin, Ming-Wei Chang, Kenton Lee, and Kristina Toutanova. 2018. BERT: Pre-training of deep bidirectional transformers for language understanding. arXiv preprint arXiv:1810.04805 (2018).","journal-title":"arXiv preprint arXiv:1810.04805"},{"key":"e_1_3_2_17_2","doi-asserted-by":"publisher","DOI":"10.18653\/v1\/P18-3021"},{"key":"e_1_3_2_18_2","first-page":"2337","volume-title":"International Conference on Machine Learning","author":"Gong Linyuan","year":"2019","unstructured":"Linyuan Gong, Di He, Zhuohan Li, Tao Qin, Liwei Wang, and Tieyan Liu. 2019. Efficient training of BERT by progressively stacking. In International Conference on Machine Learning. PMLR, 2337\u20132346."},{"key":"e_1_3_2_19_2","doi-asserted-by":"publisher","DOI":"10.18653\/v1\/D19-5522"},{"key":"e_1_3_2_20_2","doi-asserted-by":"publisher","DOI":"10.1109\/ICCITECHN.2018.8631974"},{"key":"e_1_3_2_21_2","doi-asserted-by":"publisher","DOI":"10.3390\/app14062541"},{"issue":"11","key":"e_1_3_2_22_2","article-title":"Checking the correctness of Bangla words using n-gram","volume":"89","author":"Khan Nur Hossain","year":"2014","unstructured":"Nur Hossain Khan, Gonesh Chandra Saha, Bappa Sarker, and Md Habibur Rahman. 2014. Checking the correctness of Bangla words using n-gram. International Journal of Computer Application 89, 11 (2014).","journal-title":"International Journal of Computer Application"},{"key":"e_1_3_2_23_2","article-title":"RoBERTa: A robustly optimized BERT pretraining approach","author":"Liu Yinhan","year":"2019","unstructured":"Yinhan Liu, Myle Ott, Naman Goyal, Jingfei Du, Mandar Joshi, Danqi Chen, Omer Levy, Mike Lewis, Luke Zettlemoyer, and Veselin Stoyanov. 2019. RoBERTa: A robustly optimized BERT pretraining approach. arXiv preprint arXiv:1907.11692 (2019).","journal-title":"arXiv preprint arXiv:1907.11692"},{"key":"e_1_3_2_24_2","doi-asserted-by":"publisher","DOI":"10.1109\/ICIVPR.2017.7890878"},{"key":"e_1_3_2_25_2","doi-asserted-by":"publisher","DOI":"10.21928\/uhdjst.v7n1y2023.pp43-52"},{"key":"e_1_3_2_26_2","doi-asserted-by":"publisher","DOI":"10.1016\/j.eswa.2023.121351"},{"key":"e_1_3_2_27_2","article-title":"Deep contextualized word representations. arXiv 2018","volume":"12","author":"Peters M. E.","year":"1802","unstructured":"M. E. Peters, M. Neumann, M. Iyyer, M. Gardner, C. Clark, K. Lee, and L. Zettlemoyer. 1802. Deep contextualized word representations. arXiv 2018. arXiv preprint arXiv:1802.05365 12 (1802).","journal-title":"arXiv preprint arXiv:1802.05365"},{"key":"e_1_3_2_28_2","unstructured":"Alec Radford Karthik Narasimhan Tim Salimans Ilya Sutskever et\u00a0al. 2018. Improving language understanding by generative pre-training. (2018)."},{"key":"e_1_3_2_29_2","article-title":"BSpell: A CNN-blended BERT based Bengali spell checker","author":"Rahman Chowdhury Rafeed","year":"2022","unstructured":"Chowdhury Rafeed Rahman, M. D. Rahman, Samiha Zakir, Mohammad Rafsan, and Mohammed Eunus Ali. 2022. BSpell: A CNN-blended BERT based Bengali spell checker. arXiv preprint arXiv:2208.09709 (2022).","journal-title":"arXiv preprint arXiv:2208.09709"},{"key":"e_1_3_2_30_2","first-page":"216","volume-title":"Proceedings of the 3rd Workshop on Asian Translation (WAT2016)","author":"Sen Sukanta","year":"2016","unstructured":"Sukanta Sen, Debajyoty Banik, Asif Ekbal, and Pushpak Bhattacharyya. 2016. IITP English-Hindi machine translation system at WAT 2016. In Proceedings of the 3rd Workshop on Asian Translation (WAT2016). The COLING 2016 Organizing Committee, Osaka, Japan, 216\u2013222. https:\/\/aclanthology.org\/W16-4622"},{"key":"e_1_3_2_31_2","doi-asserted-by":"publisher","DOI":"10.1109\/TENSYMP50017.2020.9230838"},{"key":"e_1_3_2_32_2","doi-asserted-by":"publisher","DOI":"10.1109\/NLPKE.2005.1598827"},{"key":"e_1_3_2_33_2","article-title":"Attention is all you need","volume":"30","author":"Vaswani Ashish","year":"2017","unstructured":"Ashish Vaswani, Noam Shazeer, Niki Parmar, Jakob Uszkoreit, Llion Jones, Aidan N. Gomez, \u0141ukasz Kaiser, and Illia Polosukhin. 2017. Attention is all you need. Advances in Neural Information Processing Systems 30 (2017).","journal-title":"Advances in Neural Information Processing Systems"},{"key":"e_1_3_2_34_2","doi-asserted-by":"publisher","DOI":"10.18653\/v1\/P19-1578"},{"key":"e_1_3_2_35_2","volume-title":"International Journal of Computational Linguistics & Chinese Language Processing, Volume 20, Number 1, June 2015\u2014Special Issue on Chinese as a Foreign Language","author":"Xiong Jinhua","year":"2015","unstructured":"Jinhua Xiong, Qiao Zhang, Shuiyuan Zhang, Jianpeng Hou, and Xueqi Cheng. 2015. HANSpeller: A unified framework for Chinese spelling correction. In International Journal of Computational Linguistics & Chinese Language Processing, Volume 20, Number 1, June 2015\u2014Special Issue on Chinese as a Foreign Language."},{"key":"e_1_3_2_36_2","article-title":"Spelling error correction with soft-masked BERT","author":"Zhang Shaohua","year":"2020","unstructured":"Shaohua Zhang, Haoran Huang, Jicong Liu, and Hang Li. 2020. Spelling error correction with soft-masked BERT. arXiv preprint arXiv:2005.07421 (2020).","journal-title":"arXiv preprint arXiv:2005.07421"}],"container-title":["ACM Transactions on Asian and Low-Resource Language Information Processing"],"original-title":[],"language":"en","link":[{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/3669941","content-type":"unspecified","content-version":"vor","intended-application":"text-mining"},{"URL":"https:\/\/dl.acm.org\/doi\/pdf\/10.1145\/3669941","content-type":"unspecified","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2025,6,19]],"date-time":"2025-06-19T00:05:44Z","timestamp":1750291544000},"score":1,"resource":{"primary":{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/3669941"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2024,8,8]]},"references-count":35,"journal-issue":{"issue":"8","published-print":{"date-parts":[[2024,8,31]]}},"alternative-id":["10.1145\/3669941"],"URL":"https:\/\/doi.org\/10.1145\/3669941","relation":{},"ISSN":["2375-4699","2375-4702"],"issn-type":[{"value":"2375-4699","type":"print"},{"value":"2375-4702","type":"electronic"}],"subject":[],"published":{"date-parts":[[2024,8,8]]},"assertion":[{"value":"2023-10-10","order":0,"name":"received","label":"Received","group":{"name":"publication_history","label":"Publication History"}},{"value":"2024-05-24","order":2,"name":"accepted","label":"Accepted","group":{"name":"publication_history","label":"Publication History"}},{"value":"2024-08-08","order":3,"name":"published","label":"Published","group":{"name":"publication_history","label":"Publication History"}}]}}