{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2026,7,24]],"date-time":"2026-07-24T23:19:11Z","timestamp":1784935151367,"version":"3.55.0"},"reference-count":45,"publisher":"Wiley","issue":"1","license":[{"start":{"date-parts":[[2021,9,21]],"date-time":"2021-09-21T00:00:00Z","timestamp":1632182400000},"content-version":"vor","delay-in-days":263,"URL":"http:\/\/creativecommons.org\/licenses\/by\/4.0\/"}],"funder":[{"DOI":"10.13039\/501100006261","name":"Taif University","doi-asserted-by":"publisher","award":["TURSP-2020\/239"],"award-info":[{"award-number":["TURSP-2020\/239"]}],"id":[{"id":"10.13039\/501100006261","id-type":"DOI","asserted-by":"publisher"}]}],"content-domain":{"domain":["onlinelibrary.wiley.com"],"crossmark-restriction":true},"short-container-title":["Computational Intelligence and Neuroscience"],"published-print":{"date-parts":[[2021,1]]},"abstract":"<jats:p>The process of detecting language from an audio clip by an unknown speaker, regardless of gender, manner of speaking, and distinct age speaker, is defined as spoken language identification (SLID). The considerable task is to recognize the features that can distinguish between languages clearly and efficiently. The model uses audio files and converts those files into spectrogram images. It applies the convolutional neural network (CNN) to bring out main attributes or features to detect output easily. The main objective is to detect languages out of English, French, Spanish, and German, Estonian, Tamil, Mandarin, Turkish, Chinese, Arabic, Hindi, Indonesian, Portuguese, Japanese, Latin, Dutch, Portuguese, Pushto, Romanian, Korean, Russian, Swedish, Tamil, Thai, and Urdu. An experiment was conducted on different audio files using the Kaggle dataset named spoken language identification. These audio files are comprised of utterances, each of them spanning over a fixed duration of 10 seconds. The whole dataset is split into training and test sets. Preparatory results give an overall accuracy of 98%. Extensive and accurate testing show an overall accuracy of 88%.<\/jats:p>","DOI":"10.1155\/2021\/5123671","type":"journal-article","created":{"date-parts":[[2021,9,22]],"date-time":"2021-09-22T03:20:23Z","timestamp":1632280823000},"update-policy":"https:\/\/doi.org\/10.1002\/crossmark_policy","source":"Crossref","is-referenced-by-count":68,"title":["Spoken Language Identification Using Deep Learning"],"prefix":"10.1155","volume":"2021","author":[{"ORCID":"https:\/\/orcid.org\/0000-0003-3826-8857","authenticated-orcid":false,"given":"Gundeep","family":"Singh","sequence":"first","affiliation":[],"role":[{"vocabulary":"crossref","role":"author"}]},{"ORCID":"https:\/\/orcid.org\/0000-0002-6694-3365","authenticated-orcid":false,"given":"Sahil","family":"Sharma","sequence":"additional","affiliation":[],"role":[{"vocabulary":"crossref","role":"author"}]},{"ORCID":"https:\/\/orcid.org\/0000-0002-3460-6989","authenticated-orcid":false,"given":"Vijay","family":"Kumar","sequence":"additional","affiliation":[],"role":[{"vocabulary":"crossref","role":"author"}]},{"ORCID":"https:\/\/orcid.org\/0000-0001-6259-2046","authenticated-orcid":false,"given":"Manjit","family":"Kaur","sequence":"additional","affiliation":[],"role":[{"vocabulary":"crossref","role":"author"}]},{"ORCID":"https:\/\/orcid.org\/0000-0003-2417-4374","authenticated-orcid":false,"given":"Mohammed","family":"Baz","sequence":"additional","affiliation":[],"role":[{"vocabulary":"crossref","role":"author"}]},{"ORCID":"https:\/\/orcid.org\/0000-0001-6019-7245","authenticated-orcid":false,"given":"Mehedi","family":"Masud","sequence":"additional","affiliation":[],"role":[{"vocabulary":"crossref","role":"author"}]}],"member":"311","published-online":{"date-parts":[[2021,9,21]]},"reference":[{"key":"e_1_2_11_1_2","doi-asserted-by":"publisher","DOI":"10.1109\/79.317925"},{"key":"e_1_2_11_2_2","unstructured":"MontavonG. Deep learning for spoken language identification Proceedings of the NIPS Workshop on Deep Learning for Speech Recognition and Related Applications December 2009 Vancouver Canada 1\u20134."},{"key":"e_1_2_11_3_2","unstructured":"KumarP. BiswasA. MishraA. N. andChandraM. Spoken language identification using hybrid feature extraction methods 2010 1 no. 2 11\u201315 https:\/\/arxiv.org\/abs\/1003.5623."},{"key":"e_1_2_11_4_2","unstructured":"SrivastavaB. M. L. VydanaH. K. VuppalaA. K. andShrivastavaM. A language model based approach towards large scale and lightweight language identification systems 2015 4\u20137 https:\/\/arxiv.org\/abs\/1510.03602."},{"key":"e_1_2_11_5_2","doi-asserted-by":"crossref","unstructured":"ShenP. LuX. LiS. andKawaiH. Conditional generative adversarial nets classifier for spoken language identification Proceedings of the Annual Conference of the International Speech Communication Association August 2017 Stockholm Sweden INTERSPEECH 2814\u20132818 https:\/\/doi.org\/10.21437\/Interspeech.2017-553 2-s2.0-85039152451.","DOI":"10.21437\/Interspeech.2017-553"},{"key":"e_1_2_11_6_2","doi-asserted-by":"crossref","unstructured":"ZampieriM. CiobanuA. M. andDinuL. P. Native language identification on text and speech 2017 https:\/\/arxiv.org\/abs\/1707.07182 https:\/\/doi.org\/10.18653\/v1\/w17-5045.","DOI":"10.18653\/v1\/W17-5045"},{"key":"e_1_2_11_7_2","doi-asserted-by":"crossref","unstructured":"CampbellW. M. RichardsonF. andReynoldsD. A. Language recognition with word lattices and support vector machines 4 Proceedings of the 2007 IEEE International Conference on Acoustics Speech and Signal Processing\u2014ICASSP \u201907 April 2007 Honolulu HI USA no. 2 989\u2013992 https:\/\/doi.org\/10.1109\/ICASSP.2007.367238 2-s2.0-34547544522.","DOI":"10.1109\/ICASSP.2007.367238"},{"key":"e_1_2_11_8_2","unstructured":"RevayS. TeschkeM. andNovetta Multi-class language identification using deep learning on spectral images of audio signals 2019 1\u20137 https:\/\/arxiv.org\/abs\/1905.04348."},{"key":"e_1_2_11_9_2","doi-asserted-by":"publisher","DOI":"10.1016\/j.comcom.2020.01.050"},{"key":"e_1_2_11_10_2","doi-asserted-by":"publisher","DOI":"10.1504\/ijiei.2020.10034278"},{"key":"e_1_2_11_11_2","first-page":"97","volume-title":"Handbook of Research on Deep Learning Innovations and Trends","author":"Sharma S.","year":"2019"},{"key":"e_1_2_11_12_2","doi-asserted-by":"publisher","DOI":"10.1007\/s12652-020-02386-0"},{"key":"e_1_2_11_13_2","doi-asserted-by":"publisher","DOI":"10.1007\/s11045-020-00739-8"},{"key":"e_1_2_11_14_2","doi-asserted-by":"publisher","DOI":"10.1007\/978-3-030-34255-5_17"},{"key":"e_1_2_11_15_2","unstructured":"Lopez-morenoI. Gonzalez-dominguezJ. PlchotO. MartinezD. Gonzalez-rodriguezJ. andMorenoP. Google inc . New York USA ATVS-biometric recognition group universidad autonoma de madrid Spain brno university of technology czech republic aragon institute for engineering research ( I3A ) university of Zaragoza Spain 2014 0\u20134."},{"key":"e_1_2_11_16_2","unstructured":"van der MerweR. Triplet entropy loss: improving the generalisation of short speech language identification systems 2020 https:\/\/arxiv.org\/abs\/2012.03775."},{"key":"e_1_2_11_17_2","unstructured":"LuX. ShenP. TsaoY. andKawaiH. Unsupervised neural adaptation model based on optimal transport for spoken language identification 2020 no. 19 https:\/\/arxiv.org\/abs\/2012.13152."},{"key":"e_1_2_11_18_2","unstructured":"AbdullahB. M. KuderaJ. AvgustinovaT. M\u00f6biusB. andKlakowD. Rediscovering the slavic continuum in representations emerging from neural models of spoken language identification 2020 https:\/\/arxiv.org\/abs\/2010.11973."},{"key":"e_1_2_11_19_2","unstructured":"RanganP. TekiS. andMisraH. Exploiting spectral augmentation for code-switched spoken language identification 2020 no. 2 https:\/\/arxiv.org\/abs\/2010.07130."},{"key":"e_1_2_11_20_2","doi-asserted-by":"crossref","unstructured":"VermaM.andBuduruA. B. Fine-grained language identification with multilingual CapsNet model Proceedings of the\u20142020 IEEE 6th International Conference on Multimedia Big Data (BigMM) September 2020 New Delhi India 94\u2013102 https:\/\/doi.org\/10.1109\/BigMM50055.2020.00023.","DOI":"10.1109\/BigMM50055.2020.00023"},{"key":"e_1_2_11_21_2","doi-asserted-by":"crossref","unstructured":"AnsariM. Z. AhmadT. andFatimaA. Feature selection on noisy twitter short text messages for language identification 2020 no. 1 https:\/\/arxiv.org\/abs\/2007.05727 https:\/\/doi.org\/10.35940\/ijrte.d4360.118419.","DOI":"10.35940\/ijrte.D4360.118419"},{"key":"e_1_2_11_22_2","unstructured":"PunjabiS. ArsikereH. RaeesyZ. ChandakC. BhaveN. BansalA. MullerM. MurilloS. RastrowA. GarimellaS. MaasR. HansM. MouchtarisA. andKunzmannS. Streaming end-to-end bilingual ASR systems with joint language identification 2020 https:\/\/arxiv.org\/abs\/2007.03900."},{"key":"e_1_2_11_23_2","unstructured":"ChandakC. RaeesyZ. RastrowA. LiuY. HuangX. WangS. JooD. K. andMassR. Streaming language identification using combination of acoustic representations and ASR hypotheses 2020 https:\/\/arxiv.org\/abs\/2006.00703."},{"key":"e_1_2_11_24_2","unstructured":"WangS. WanL. YuY. andMorenoI. L. Signal combination for language identification 2019 https:\/\/arxiv.org\/abs\/1910.09687."},{"key":"e_1_2_11_25_2","doi-asserted-by":"crossref","unstructured":"TitusA. SilovskyJ. ChenN. HsiaoR. YoungM. andGhoshalA. Improving language identification for multilingual speakers Proceedings of the ICASSP 2020\u20142020 IEEE International Conference on Acoustics Speech and Signal Processing (ICASSP) May 2020 Barcelona Spain 8284\u20138288 https:\/\/doi.org\/10.1109\/ICASSP40776.2020.9053057.","DOI":"10.1109\/ICASSP40776.2020.9053057"},{"key":"e_1_2_11_26_2","doi-asserted-by":"publisher","DOI":"10.11591\/eei.v10i4.2893"},{"key":"e_1_2_11_27_2","doi-asserted-by":"crossref","unstructured":"SaleskyE. AbdullahB. M. MielkeS. KlyachkoE. SerikovO. PontiE. M. KumarR. CotterellR. andVylomovaE. SIGTYP 2021 shared task: robust spoken language identification Proceedings of the Third Workshop on Computational Typology and Multilingual NLP Association for Computational Linguistics 2021 Mexico city Mexico 122\u2013129 https:\/\/doi.org\/10.18653\/v1\/2021.sigtyp-1.11.","DOI":"10.18653\/v1\/2021.sigtyp-1.11"},{"key":"e_1_2_11_28_2","doi-asserted-by":"crossref","unstructured":"BedyakinR. Low-resource spoken language identification using self-attentive pooling and deep 1D time-channel separable convolutions 2021 http:\/\/arxiv.org\/abs\/2106.00052v1.","DOI":"10.28995\/2075-7182-2021-20-1012-1020"},{"key":"e_1_2_11_29_2","doi-asserted-by":"crossref","unstructured":"ScherbakovA. WhittleL. KumarR. SinghS. ColemanM. andVylomovaE. Anlirika: an LSTM\u2013CNN flow twister for spoken language identification Proceedings of the Third Workshop on Computational Typology and Multilingual NLP Association for Computational Linguistics 2021 Mexico city Mexico 145\u2013148 https:\/\/doi.org\/10.18653\/v1\/2021.sigtyp-1.14.","DOI":"10.18653\/v1\/2021.sigtyp-1.14"},{"key":"e_1_2_11_30_2","unstructured":"Spoken Language Identification | Kaggle 2021 https:\/\/www.kaggle.com\/toponowicz\/spoken-language-identification."},{"key":"e_1_2_11_31_2","unstructured":"Language Identification dataset | Kaggle 2021 https:\/\/www.kaggle.com\/zarajamshaid\/language-identification-datasst."},{"key":"e_1_2_11_32_2","unstructured":"Common Voice | Kaggle 2021 https:\/\/www.kaggle.com\/mozillaorg\/common-voice."},{"key":"e_1_2_11_33_2","unstructured":"Common Voice 2021 https:\/\/commonvoice.mozilla.org\/en."},{"key":"e_1_2_11_34_2","doi-asserted-by":"crossref","unstructured":"RanasingheT.andZampieriM. Multilingual offensive language identification with cross-lingual embeddings 2020 https:\/\/arxiv.org\/abs\/2010.05324 https:\/\/doi.org\/10.18653\/v1\/2020.emnlp-main.470.","DOI":"10.18653\/v1\/2020.emnlp-main.470"},{"key":"e_1_2_11_35_2","unstructured":"OrasanC. Aggressive language identification using word embeddings and sentiment features Proceedings of the First Workshop on Trolling Aggression and Cyberbullying ({TRAC}-2018) August 2018 Santa Fe NM USA 113\u2013119 https:\/\/www.aclweb.org\/anthology\/W18-4414."},{"key":"e_1_2_11_36_2","doi-asserted-by":"publisher","DOI":"10.1007\/978-3-642-12275-0_59"},{"key":"e_1_2_11_37_2","unstructured":"BaldwinT.andLuiM. Language identification: the long and the short of the matter Proceedings of the NAACL HLT 2010\u2014Human Language Technologies: the 2010 Annual Conference of the North American Chapter of the Association for Computational Linguistics Proceedings of the Main Conference June 2010 Los Angeles CA USA 229\u2013237."},{"key":"e_1_2_11_38_2","first-page":"25","article-title":"Langid.py: an off-the-shelf Language identification tool","author":"Lui M.","year":"2012","journal-title":"Aclweb.Org"},{"key":"e_1_2_11_39_2","doi-asserted-by":"publisher","DOI":"10.1109\/ACCESS.2021.3101142"},{"key":"e_1_2_11_40_2","doi-asserted-by":"publisher","DOI":"10.37965\/jait.2020.0051"},{"key":"e_1_2_11_41_2","doi-asserted-by":"publisher","DOI":"10.37965\/jait.2020.0037"},{"key":"e_1_2_11_42_2","doi-asserted-by":"publisher","DOI":"10.1049\/trit.2019.0051"},{"key":"e_1_2_11_43_2","doi-asserted-by":"publisher","DOI":"10.1049\/trit.2018.1006"},{"key":"e_1_2_11_44_2","doi-asserted-by":"publisher","DOI":"10.3390\/s21030748"},{"key":"e_1_2_11_45_2","doi-asserted-by":"publisher","DOI":"10.37965\/jait.2021.0017"}],"container-title":["Computational Intelligence and Neuroscience"],"original-title":[],"language":"en","link":[{"URL":"http:\/\/downloads.hindawi.com\/journals\/cin\/2021\/5123671.pdf","content-type":"application\/pdf","content-version":"vor","intended-application":"text-mining"},{"URL":"http:\/\/downloads.hindawi.com\/journals\/cin\/2021\/5123671.xml","content-type":"application\/xml","content-version":"vor","intended-application":"text-mining"},{"URL":"https:\/\/onlinelibrary.wiley.com\/doi\/pdf\/10.1155\/2021\/5123671","content-type":"unspecified","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2024,8,6]],"date-time":"2024-08-06T11:46:50Z","timestamp":1722944810000},"score":1,"resource":{"primary":{"URL":"https:\/\/onlinelibrary.wiley.com\/doi\/10.1155\/2021\/5123671"}},"subtitle":[],"editor":[{"given":"Ciro","family":"Castiello","sequence":"additional","affiliation":[],"role":[{"vocabulary":"crossref","role":"editor"}]}],"short-title":[],"issued":{"date-parts":[[2021,1]]},"references-count":45,"journal-issue":{"issue":"1","published-print":{"date-parts":[[2021,1]]}},"alternative-id":["10.1155\/2021\/5123671"],"URL":"https:\/\/doi.org\/10.1155\/2021\/5123671","archive":["Portico"],"relation":{},"ISSN":["1687-5265","1687-5273"],"issn-type":[{"value":"1687-5265","type":"print"},{"value":"1687-5273","type":"electronic"}],"subject":[],"published":{"date-parts":[[2021,1]]},"assertion":[{"value":"2021-07-06","order":0,"name":"received","label":"Received","group":{"name":"publication_history","label":"Publication History"}},{"value":"2021-09-03","order":2,"name":"accepted","label":"Accepted","group":{"name":"publication_history","label":"Publication History"}},{"value":"2021-09-21","order":3,"name":"published","label":"Published","group":{"name":"publication_history","label":"Publication History"}}],"article-number":"5123671"}}