{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2026,5,2]],"date-time":"2026-05-02T06:29:20Z","timestamp":1777703360523,"version":"3.51.4"},"reference-count":20,"publisher":"SAGE Publications","issue":"3","license":[{"start":{"date-parts":[[2018,3,22]],"date-time":"2018-03-22T00:00:00Z","timestamp":1521676800000},"content-version":"tdm","delay-in-days":0,"URL":"https:\/\/journals.sagepub.com\/page\/policies\/text-and-data-mining-license"}],"content-domain":{"domain":["journals.sagepub.com"],"crossmark-restriction":true},"short-container-title":["Journal of Intelligent &amp; Fuzzy Systems"],"published-print":{"date-parts":[[2018,3,22]]},"abstract":"<jats:p>The present work describes different research techniques for collecting and organizing speech database in different scenario at the institute and successfully structuring the text independent speaker identification database in Indian context. In order to get the Multi-Scenario dataset, each speaker performed multiple sessions recording in reading style with English and Hindi language with same passages but under different conditions. This work analyzed different scenario affecting the performance of speaker recognition system when tested under dissimilar training conditions. Here four different scenarios are considered; sensor and environment, language, aging and health. To study the effect of sensor, language and environment on the performance of ASR system a database of 200 speaker was created. Under different environmental conditions, four different types of sensors in parallel configuration were used to study the sensor mismatch conditions over testing and training phase. The database contains speech samples of the individual in English and Hindi in read speech styles under two environment i.e. a controlled recording chamber and library. To study the aging effect, an aging NSIT speaker database (AG-NSIT-SD) of 53 famous personalities was collected from online source varying over a period of 10\u201320 years. Further to study the effect of health, a cough and cold NSIT speaker database (CC-NSIT-SD) of 38 speakers was also collected to study the performance of system. Apart from this, the effect of different noise types on the speaker identification was also studied on different sensors.<\/jats:p>","DOI":"10.3233\/jifs-169433","type":"journal-article","created":{"date-parts":[[2018,3,23]],"date-time":"2018-03-23T12:21:34Z","timestamp":1521807694000},"page":"1385-1392","update-policy":"https:\/\/doi.org\/10.1177\/sage-journals-update-policy","source":"Crossref","is-referenced-by-count":3,"title":["Multi-scenario dataset for speaker recognition"],"prefix":"10.1177","volume":"34","author":[{"given":"Smriti","family":"Srivastava","sequence":"first","affiliation":[{"name":"Netaji Subhas Institute of Technology, Delhi University, Dwarka, Delhi, India"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"family":"Gopal","sequence":"additional","affiliation":[{"name":"Bharati Vidyapeeth\u2019s College of Engineering, Guru Gobind Singh Indraprastha University, Delhi, India"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Saurabh","family":"Bhardwaj","sequence":"additional","affiliation":[{"name":"Thapar University, Patiala, punjab, India"}],"role":[{"role":"author","vocabulary":"crossref"}]}],"member":"179","published-online":{"date-parts":[[2018,3,22]]},"reference":[{"key":"e_1_3_1_2_2","first-page":"1651","volume-title":"Speaker recognitionidentifying people by their voices","volume":"73","author":"Doddington G.R.","year":"1985","unstructured":"DoddingtonG.R., Speaker recognitionidentifying people by their voices, Proceedings of the IEEE73(11) (1985), 1651\u20131664."},{"key":"e_1_3_1_3_2","first-page":"829","volume-title":"Corpora for the evaluation of speaker recognition systems","author":"Campbell J.P.","year":"1999","unstructured":"CampbellJ.P.Jr and ReynoldsD.A., Corpora for the evaluation of speaker recognition systems, In: Acoustics, Speech, and Signal Processing, 1999. Proceedings., 1999 IEEE International Conference on, IEEE, (1999), 829\u2013832."},{"key":"e_1_3_1_4_2","first-page":"4072","volume-title":"An overview of automatic speaker recognition","author":"Reynolds D.","year":"2002","unstructured":"ReynoldsD., An overview of automatic speaker recognition, In: Proceedings of the International Conference on Acoustics, Speech and Signal Processing (ICASSP), Springer (2002), 4072\u20134075."},{"key":"e_1_3_1_5_2","doi-asserted-by":"crossref","first-page":"424","DOI":"10.1007\/978-3-319-10816-2_51","volume-title":"Text, Speech and Dialogue","author":"Neuberger T.","year":"2014","unstructured":"NeubergerT., GyarmathyD., GracziT.E., HorvathV., GosyM. and BekeA., Development \u00c2t\u2019 of a large spontaneous speech database of agglutinative hungarian language, In: Text, Speech and Dialogue, Springer (2014), 424\u2013431."},{"key":"e_1_3_1_6_2","unstructured":"GarofoloJ.S. LamelL.F. FisherW.M. FiscusJ.G. PallettD.S. and DahlgrenN.L. TIMIT acoustic-phonetic continuous speech corpus Philadelphia: Linguistic Data Consortium 1993."},{"key":"e_1_3_1_7_2","unstructured":"FengL. and HansenL.K. A new database for speaker recognition. Tech. rep 2005."},{"key":"e_1_3_1_8_2","first-page":"109","volume-title":"Ntimit: A phonetically balanced, continuous speech, telephone bandwidth speech database","author":"Jankowski C.","year":"1990","unstructured":"JankowskiC., KalyanswamyA., BassonS. and SpitzJ., Ntimit: A phonetically balanced, continuous speech, telephone bandwidth speech database, In: Acoustics, Speech, and Signal Processing, 1990. ICASSP-90., 1990 International Conference on. pp. 109\u2013112. IEEE (1990)."},{"key":"e_1_3_1_9_2","first-page":"113","volume-title":"The effects of handset variability on speaker recognition performance: Experiments on the switchboard corpus","author":"Reynolds D.A.","year":"1996","unstructured":"ReynoldsD.A., The effects of handset variability on speaker recognition performance: Experiments on the switchboard corpus, In: Acoustics, Speech, and Signal Processing, 1996. ICASSP-96. Conference Proceedings., 1996 IEEE International Conference on, IEEE (1996), 113\u2013116."},{"key":"e_1_3_1_10_2","first-page":"96","volume-title":"The atis spoken language systems pilot corpus","author":"Hemphill C.T.","year":"1990","unstructured":"HemphillC.T., GodfreyJ.J. and DoddingtonG.R., The atis spoken language systems pilot corpus, In: Proceedings of the DARPA speech and natural language workshop (1990), 96\u2013101."},{"key":"e_1_3_1_11_2","volume-title":"Reducing inter-session variability with transitional spectral information","author":"Wildermoth B.R.","year":"2001","unstructured":"WildermothB.R. and PaliwalK.K., Reducing inter-session variability with transitional spectral information, In: Proc of Microelectronic Engineering Research Conference (2001)."},{"key":"e_1_3_1_12_2","article-title":"The aurora experimental framework for the performance evaluation of speech recognition systems under noisy conditions","author":"Hirsch H.G.","year":"2000","unstructured":"HirschH.G. and PearceD., The aurora experimental framework for the performance evaluation of speech recognition systems under noisy conditions In: ASR2000-Automatic Seech Recognition: Challenges for the new Millenium ISCA Tutorial and Research Workshop (ITRW) (2000).","journal-title":"ASR2000-Automatic Seech Recognition: Challenges for the new Millenium ISCA Tutorial and Research Workshop (ITRW)"},{"key":"e_1_3_1_13_2","unstructured":"CholletG. CochardJ.L. ConstantinescuA. JabouletC. and LanglaisP. Swiss french polyphone and polyvar: Telephone speech databases to model inter-and intra-speaker variability IDIAP-RR (1996)."},{"key":"e_1_3_1_14_2","first-page":"61","article-title":"Speechdat-speech databases for creation of voice driven teleservices","volume":"4","author":"Elenius K.","year":"1997","unstructured":"EleniusK. and LindbergJ., Speechdat-speech databases for creation of voice driven teleservices, Proc of Fonetik-97, Dept of Phonetics, Phonum4 (1997), 61\u201364.","journal-title":"Proc of Fonetik-97, Dept of Phonetics, Phonum"},{"issue":"2","key":"e_1_3_1_15_2","doi-asserted-by":"crossref","first-page":"265","DOI":"10.1016\/S0167-6393(99)00082-5","article-title":"Polycost: A telephone-speech database for speaker recognition","volume":"31","author":"Hennebert J.","year":"2000","unstructured":"HennebertJ., MelinH., PetrovskaD. and GenoudD., Polycost: A telephone-speech database for speaker recognition, Speech Communication31(2) (2000), 265\u2013270.","journal-title":"Speech Communication"},{"key":"e_1_3_1_16_2","first-page":"1902","volume-title":"The siva speech database for speaker verification: Description and evaluation","author":"Falcone M.","year":"1996","unstructured":"FalconeM. and GalloA., The siva speech database for speaker verification: Description and evaluation, In: Spoken Language, 1996. ICSLP 96. Proceedings., Fourth International Conference on, IEEE (1996), 1902\u20131905."},{"key":"e_1_3_1_17_2","unstructured":"HigginsA. and VermilyeaD. King speaker verification lDC catalog number LDC95S22 1995."},{"key":"e_1_3_1_18_2","unstructured":"https:\/\/www.youtube.com\/"},{"key":"e_1_3_1_19_2","unstructured":"http:\/\/www.dailymotion.com\/in\/"},{"key":"e_1_3_1_20_2","unstructured":"http:\/\/www.nist.gov\/index.html"},{"key":"e_1_3_1_21_2","doi-asserted-by":"publisher","DOI":"10.1109\/TSMCB.2012.2223461"}],"container-title":["Journal of Intelligent &amp; Fuzzy Systems"],"original-title":[],"language":"en","link":[{"URL":"https:\/\/journals.sagepub.com\/doi\/pdf\/10.3233\/JIFS-169433","content-type":"application\/pdf","content-version":"vor","intended-application":"text-mining"},{"URL":"https:\/\/journals.sagepub.com\/doi\/full-xml\/10.3233\/JIFS-169433","content-type":"application\/xml","content-version":"vor","intended-application":"text-mining"},{"URL":"https:\/\/journals.sagepub.com\/doi\/pdf\/10.3233\/JIFS-169433","content-type":"unspecified","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2026,4,29]],"date-time":"2026-04-29T09:38:30Z","timestamp":1777455510000},"score":1,"resource":{"primary":{"URL":"https:\/\/journals.sagepub.com\/doi\/10.3233\/JIFS-169433"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2018,3,22]]},"references-count":20,"journal-issue":{"issue":"3","published-print":{"date-parts":[[2018,3,22]]}},"alternative-id":["10.3233\/JIFS-169433"],"URL":"https:\/\/doi.org\/10.3233\/jifs-169433","relation":{},"ISSN":["1064-1246","1875-8967"],"issn-type":[{"value":"1064-1246","type":"print"},{"value":"1875-8967","type":"electronic"}],"subject":[],"published":{"date-parts":[[2018,3,22]]}}}