{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2025,6,18]],"date-time":"2025-06-18T04:13:11Z","timestamp":1750219991527,"version":"3.41.0"},"reference-count":38,"publisher":"Association for Computing Machinery (ACM)","issue":"1","license":[{"start":{"date-parts":[[2024,1,15]],"date-time":"2024-01-15T00:00:00Z","timestamp":1705276800000},"content-version":"vor","delay-in-days":0,"URL":"https:\/\/www.acm.org\/publications\/policies\/copyright_policy#Background"}],"content-domain":{"domain":["dl.acm.org"],"crossmark-restriction":true},"short-container-title":["ACM Trans. Asian Low-Resour. Lang. Inf. Process."],"published-print":{"date-parts":[[2024,1,31]]},"abstract":"<jats:p>One common constraint in the practical application of speech recognition is Code Switching. The issue of code-switched languages is especially aggravated in the context of Indian languages \u2013 since most massively multilingual models are trained on corpora that are not representative of the diverse set of Indian languages. An associated constraint with such systems is the privacy-intrusive nature of the applications that aim to collate such representative data. To collectively mitigate both problems, this work presents CodeFed: A federated learning-based code-switching detection model that can be deployed to collaboratively be trained by leveraging private data from multiple users, without compromising their privacy. Using a representative low-resource Indic dataset, we demonstrate the superior performance of a collaboratively trained global model that is trained using federated learning on three low-resource Indic languages \u2013 Gujarati, Tamil and Telugu and draw a comparison of the model with respect to the most current work in the field. Finally, to evaluate the practical realizability of the proposed system, CodeFed also discusses the system overview of the label generation architecture which may accompany CodeFed\u2019s possible real-time deployment.<\/jats:p>","DOI":"10.1145\/3571732","type":"journal-article","created":{"date-parts":[[2022,11,17]],"date-time":"2022-11-17T15:05:16Z","timestamp":1668697516000},"page":"1-14","update-policy":"https:\/\/doi.org\/10.1145\/crossmark-policy","source":"Crossref","is-referenced-by-count":1,"title":["CodeFed: Federated Speech Recognition for Low-Resource Code-Switching Detection"],"prefix":"10.1145","volume":"23","author":[{"ORCID":"https:\/\/orcid.org\/0000-0002-4534-1621","authenticated-orcid":false,"given":"Chetan","family":"Madan","sequence":"first","affiliation":[{"name":"Department of Computer Science and Engineering, Bharati Vidyapeeth\u2019s College of Engineering, New Delhi, India"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"ORCID":"https:\/\/orcid.org\/0000-0002-0852-7371","authenticated-orcid":false,"given":"Harshita","family":"Diddee","sequence":"additional","affiliation":[{"name":"Department of Computer Science and Engineering, Bharati Vidyapeeth\u2019s College of Engineering, New Delhi, India"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"ORCID":"https:\/\/orcid.org\/0000-0002-6690-8500","authenticated-orcid":false,"given":"Deepika","family":"Kumar","sequence":"additional","affiliation":[{"name":"Department of Computer Science and Engineering, Bharati Vidyapeeth\u2019s College of Engineering, New Delhi, India"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"ORCID":"https:\/\/orcid.org\/0000-0003-0490-4413","authenticated-orcid":false,"given":"Mamta","family":"Mittal","sequence":"additional","affiliation":[{"name":"Delhi Skill and Entrepreneurship University, India"}],"role":[{"role":"author","vocabulary":"crossref"}]}],"member":"320","published-online":{"date-parts":[[2024,1,15]]},"reference":[{"key":"e_1_3_1_2_2","doi-asserted-by":"publisher","DOI":"10.1109\/ICACCS48705.2020.9074205"},{"key":"e_1_3_1_3_2","first-page":"102","volume-title":"Selected Proceedings of the Second Workshop on Spanish Sociolinguistics","author":"Montes-Alcal\u00e1 Cecilia","year":"2005","unstructured":"Cecilia Montes-Alcal\u00e1. 2005. Dear amigo: Exploring code-switching in personal letters. In Selected Proceedings of the Second Workshop on Spanish Sociolinguistics. Cascadilla Proceedings Project Somerville, MA, 102\u2013108."},{"key":"e_1_3_1_4_2","doi-asserted-by":"publisher","DOI":"10.1016\/j.csl.2021.101278"},{"key":"e_1_3_1_5_2","article-title":"CALCS 2021 shared task: Machine translation for code-switched data","author":"Chen Shuguang","year":"2022","unstructured":"Shuguang Chen, Gustavo Aguilar, Anirudh Srinivasan, Mona Diab, and Thamar Solorio. 2022. CALCS 2021 shared task: Machine translation for code-switched data. arXiv preprint arXiv:2202.09625 (2022).","journal-title":"arXiv preprint arXiv:2202.09625"},{"key":"e_1_3_1_6_2","doi-asserted-by":"crossref","first-page":"650","DOI":"10.18653\/v1\/2022.findings-naacl.49","volume-title":"Findings of the Association for Computational Linguistics: NAACL 2022","author":"James Jesin","year":"2022","unstructured":"Jesin James, Vithya Yogarajan, Isabella Shields, Catherine I. Watson, Peter Keegan, Keoni Mahelona, and Peter-Lucas Jones. 2022. Language models for code-switch detection of te reo M\u0101ori and English in a low-resource setting. In Findings of the Association for Computational Linguistics: NAACL 2022. 650\u2013660."},{"key":"e_1_3_1_7_2","article-title":"Low-resource machine translation for low-resource languages: Leveraging comparable data, code-switching and compute resources","author":"Kuwanto Garry","year":"2021","unstructured":"Garry Kuwanto, Afra Feyza Aky\u00fcrek, Isidora Chara Tourni, Siyang Li, and Derry Wijaya. 2021. Low-resource machine translation for low-resource languages: Leveraging comparable data, code-switching and compute resources. arXiv preprint arXiv:2103.13272 (2021).","journal-title":"arXiv preprint arXiv:2103.13272"},{"key":"e_1_3_1_8_2","unstructured":"Andrew Hard Kanishka Rao Rajiv Mathews Fran\u00e7oise Beaufays Sean Augenstein Hubert Eichner Chlo\u00e9 Kiddon and Daniel Ramage. 2018. Federated Learning for Mobile Keyboard Prediction. (112018)."},{"key":"e_1_3_1_9_2","doi-asserted-by":"publisher","DOI":"10.1109\/ICASSP39728.2021.9413397"},{"key":"e_1_3_1_10_2","doi-asserted-by":"publisher","DOI":"10.1109\/ICASSP.2019.8683546"},{"key":"e_1_3_1_11_2","doi-asserted-by":"publisher","DOI":"10.3390\/languages7030220"},{"key":"e_1_3_1_12_2","doi-asserted-by":"publisher","DOI":"10.1109\/JSAC.2019.2904348"},{"key":"e_1_3_1_13_2","article-title":"Federated learning of deep networks using model averaging","volume":"1602","author":"McMahan H. Brendan","year":"2016","unstructured":"H. Brendan McMahan, Eider Moore, Daniel Ramage, and Blaise Ag\u00fcera y Arcas. 2016. Federated learning of deep networks using model averaging. CoRR abs\/1602.05629 (2016). arXiv:1602.05629 http:\/\/arxiv.org\/abs\/1602.05629.","journal-title":"CoRR"},{"key":"e_1_3_1_14_2","doi-asserted-by":"publisher","DOI":"10.1145\/3298981"},{"key":"e_1_3_1_15_2","article-title":"Federated transfer learning with dynamic gradient aggregation","volume":"2008","author":"Dimitriadis Dimitrios","year":"2020","unstructured":"Dimitrios Dimitriadis, Ken\u2019ichi Kumatani, Robert Gmyr, Yashesh Gaur, and Sefik Emre Eskimez. 2020. Federated transfer learning with dynamic gradient aggregation. CoRR abs\/2008.02452 (2020). arXiv:2008.02452 https:\/\/arxiv.org\/abs\/2008.02452.","journal-title":"CoRR"},{"key":"e_1_3_1_16_2","doi-asserted-by":"publisher","DOI":"10.18653\/v1\/W18-6115"},{"key":"e_1_3_1_17_2","first-page":"65","volume-title":"Proceedings of the 14th International Conference on Natural Language Processing (ICON-2017)","author":"Choudhury Monojit","year":"2017","unstructured":"Monojit Choudhury, Kalika Bali, Sunayana Sitaram, and Ashutosh Baheti. 2017. Curriculum design for code-switching: Experiments with language identification and language modeling with deep neural networks. In Proceedings of the 14th International Conference on Natural Language Processing (ICON-2017). NLP Association of India, Kolkata, India, 65\u201374. https:\/\/aclanthology.org\/W17-7509."},{"key":"e_1_3_1_18_2","doi-asserted-by":"publisher","DOI":"10.1016\/j.apacoust.2019.107175"},{"key":"e_1_3_1_19_2","doi-asserted-by":"publisher","DOI":"10.18653\/v1\/D18-1347"},{"key":"e_1_3_1_20_2","article-title":"Code-switching sentence generation by generative adversarial networks and its application to data augmentation","volume":"1811","author":"Chang Ching-Ting","year":"2018","unstructured":"Ching-Ting Chang, Shun-Po Chuang, and Hung-yi Lee. 2018. Code-switching sentence generation by generative adversarial networks and its application to data augmentation. CoRR abs\/1811.02356 (2018). arXiv:1811.02356 http:\/\/arxiv.org\/abs\/1811.02356.","journal-title":"CoRR"},{"key":"e_1_3_1_21_2","doi-asserted-by":"publisher","DOI":"10.1109\/SLT.2018.8639674"},{"key":"e_1_3_1_22_2","article-title":"Semi-supervised acoustic model training for speech with code-switching","volume":"1810","author":"Yilmaz Emre","year":"2018","unstructured":"Emre Yilmaz, Mitchell McLaren, Henk van den Heuvel, and David A. van Leeuwen. 2018. Semi-supervised acoustic model training for speech with code-switching. CoRR abs\/1810.09699 (2018). arXiv:1810.09699 http:\/\/arxiv.org\/abs\/1810.09699.","journal-title":"CoRR"},{"key":"e_1_3_1_23_2","doi-asserted-by":"publisher","DOI":"10.18653\/v1\/K17-1038"},{"key":"e_1_3_1_24_2","doi-asserted-by":"publisher","DOI":"10.18653\/v1\/2020.acl-main.716"},{"key":"e_1_3_1_25_2","doi-asserted-by":"publisher","DOI":"10.48550\/ARXIV.2006.05257"},{"key":"e_1_3_1_26_2","doi-asserted-by":"publisher","DOI":"10.18653\/v1\/W18-3206"},{"key":"e_1_3_1_27_2","doi-asserted-by":"publisher","DOI":"10.18653\/v1\/N19-1201"},{"key":"e_1_3_1_28_2","first-page":"24","volume-title":"WSTCSMC 2020","author":"James Jesin","year":"2020","unstructured":"Jesin James, Vithya Yogarajan, Isabella Shields, Catherine I. Watson, Peter Keegan, Keoni Mahelona, and Peter-Lucas Jones. 2020. First workshop on speech processing for code-switching in multilingual communities: Shared task on code-switched spoken language identification. In WSTCSMC 2020. 24."},{"key":"e_1_3_1_29_2","first-page":"24","volume-title":"WSTCSMC 2020","author":"Nagarsheth J. A. C. Parav","year":"2020","unstructured":"J. A. C. Parav Nagarsheth. 2020. Language identification for codemixed Indian languages in the wild. In WSTCSMC 2020. 24."},{"key":"e_1_3_1_30_2","first-page":"53","volume-title":"WSTCSMC","author":"Patil A.","year":"2020","unstructured":"A. Patil and D. N. Krishna. 2020. Utterance-level code-switching identification using transformer network. In WSTCSMC 491 (2020), 53."},{"key":"e_1_3_1_31_2","first-page":"42","volume-title":"WSTCSMC","author":"Rallabandi Sai Krishna","year":"2020","unstructured":"Sai Krishna Rallabandi and Alan W. Black. 2020. On detecting code mixing in speech using discrete latent representations. In WSTCSMC 493 (2020), 42."},{"key":"e_1_3_1_32_2","article-title":"Improving on-device speaker verification using federated learning with privacy","author":"Granqvist Filip","year":"2020","unstructured":"Filip Granqvist, Matt Seigel, Rogier van Dalen, \u00c1ine Cahill, Stephen Shum, and Matthias Paulik. 2020. Improving on-device speaker verification using federated learning with privacy. arXiv preprint arXiv:2008.02651 (2020).","journal-title":"arXiv preprint arXiv:2008.02651"},{"key":"e_1_3_1_33_2","doi-asserted-by":"publisher","DOI":"10.1109\/WASPAA52581.2021.9632783"},{"key":"e_1_3_1_34_2","article-title":"Decoupled federated learning for ASR with non-IID data","author":"Zhu Han","year":"2022","unstructured":"Han Zhu, Jindong Wang, Gaofeng Cheng, Pengyuan Zhang, and Yonghong Yan. 2022. Decoupled federated learning for ASR with non-IID data. arXiv preprint arXiv:2206.09102 (2022).","journal-title":"arXiv preprint arXiv:2206.09102"},{"key":"e_1_3_1_35_2","article-title":"Federated learning meets natural language processing: A survey","author":"Liu Ming","year":"2021","unstructured":"Ming Liu, Stella Ho, Mengqi Wang, Longxiang Gao, Yuan Jin, and He Zhang. 2021. Federated learning meets natural language processing: A survey. arXiv preprint arXiv:2107.12603 (2021).","journal-title":"arXiv preprint arXiv:2107.12603"},{"key":"e_1_3_1_36_2","article-title":"Federated learning of n-gram language models","author":"Chen Mingqing","year":"2019","unstructured":"Mingqing Chen, Ananda Theertha Suresh, Rajiv Mathews, Adeline Wong, Cyril Allauzen, Fran\u00e7oise Beaufays, and Michael Riley. 2019. Federated learning of n-gram language models. arXiv preprint arXiv:1910.03432 (2019).","journal-title":"arXiv preprint arXiv:1910.03432"},{"key":"e_1_3_1_37_2","doi-asserted-by":"publisher","DOI":"10.1109\/ICTC52510.2021.9620836"},{"key":"e_1_3_1_38_2","first-page":"1","volume-title":"Speech Communication; 14th ITG Conference","author":"Yu Wentao","year":"2021","unstructured":"Wentao Yu, Jan Freiwald, S\u00f6ren Tewes, Fabien Huennemeyer, and Dorothea Kolossa. 2021. Federated learning in ASR: Not as easy as you think. In Speech Communication; 14th ITG Conference. VDE, 1\u20135."},{"key":"e_1_3_1_39_2","article-title":"Federated learning with dynamic transformer for text to speech","author":"Hong Zhenhou","year":"2021","unstructured":"Zhenhou Hong, Jianzong Wang, Xiaoyang Qu, Jie Liu, Chendong Zhao, and Jing Xiao. 2021. Federated learning with dynamic transformer for text to speech. arXiv preprint arXiv:2107.08795 (2021).","journal-title":"arXiv preprint arXiv:2107.08795"}],"container-title":["ACM Transactions on Asian and Low-Resource Language Information Processing"],"original-title":[],"language":"en","link":[{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/3571732","content-type":"unspecified","content-version":"vor","intended-application":"text-mining"},{"URL":"https:\/\/dl.acm.org\/doi\/pdf\/10.1145\/3571732","content-type":"unspecified","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2025,6,17]],"date-time":"2025-06-17T17:49:33Z","timestamp":1750182573000},"score":1,"resource":{"primary":{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/3571732"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2024,1,15]]},"references-count":38,"journal-issue":{"issue":"1","published-print":{"date-parts":[[2024,1,31]]}},"alternative-id":["10.1145\/3571732"],"URL":"https:\/\/doi.org\/10.1145\/3571732","relation":{},"ISSN":["2375-4699","2375-4702"],"issn-type":[{"type":"print","value":"2375-4699"},{"type":"electronic","value":"2375-4702"}],"subject":[],"published":{"date-parts":[[2024,1,15]]},"assertion":[{"value":"2022-06-27","order":0,"name":"received","label":"Received","group":{"name":"publication_history","label":"Publication History"}},{"value":"2022-11-06","order":1,"name":"accepted","label":"Accepted","group":{"name":"publication_history","label":"Publication History"}},{"value":"2024-01-15","order":2,"name":"published","label":"Published","group":{"name":"publication_history","label":"Publication History"}}]}}