{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2026,7,8]],"date-time":"2026-07-08T00:18:00Z","timestamp":1783469880240,"version":"3.55.0"},"reference-count":24,"publisher":"IGI Global Scientific Publishing","issue":"2","license":[{"start":{"date-parts":[[2019,4,1]],"date-time":"2019-04-01T00:00:00Z","timestamp":1554076800000},"content-version":"vor","delay-in-days":0,"URL":"http:\/\/creativecommons.org\/licenses\/by\/3.0\/deed.en_US"},{"start":{"date-parts":[[2019,4,1]],"date-time":"2019-04-01T00:00:00Z","timestamp":1554076800000},"content-version":"am","delay-in-days":0,"URL":"http:\/\/creativecommons.org\/licenses\/by\/3.0\/deed.en_US"},{"start":{"date-parts":[[2019,4,1]],"date-time":"2019-04-01T00:00:00Z","timestamp":1554076800000},"content-version":"tdm","delay-in-days":0,"URL":"http:\/\/creativecommons.org\/licenses\/by\/3.0\/deed.en_US"}],"content-domain":{"domain":[],"crossmark-restriction":false},"short-container-title":[],"published-print":{"date-parts":[[2019,4,1]]},"abstract":"<p>In previous studies of synthetic speech detection (SSD), the most widely used features are based on a linear power spectrum. Different from conventional methods, this article proposes a new feature extraction method for SSD from octave power spectrum which is obtained from constant-Q transform (CQT). By combining CQT, block transform (BT) and discrete cosine transform (DCT), a new feature is obtained, namely, constant-Q block coefficients (CBC). In which, CQT is used to transform speech from the time domain into the frequency domain, BT is used to segment octave power spectrum into many blocks and DCT is used to extract principal information of every block. The experimental results on ASVspoof 2015 corpus shows that CBC is superior to other front-ends features that have been benchmarked on ASVspoof 2015 evaluation set in terms of equal error rate (EER).<\/p>","DOI":"10.4018\/ijdcf.2019040105","type":"journal-article","created":{"date-parts":[[2019,2,22]],"date-time":"2019-02-22T10:12:45Z","timestamp":1550830365000},"page":"63-74","source":"Crossref","is-referenced-by-count":5,"title":["CBC-Based Synthetic Speech Detection"],"prefix":"10.4018","volume":"11","author":[{"given":"Jichen","family":"Yang","sequence":"first","affiliation":[{"name":"School of Electronic and Information Engineering, South China University of Technology, Guangzhou, China"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Qianhua","family":"He","sequence":"additional","affiliation":[{"name":"School of Electronic and Information Engineering, South China University of Technology, Guangzhou, China"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Yongjian","family":"Hu","sequence":"additional","affiliation":[{"name":"School of Electronic and Information Engineering, South China University of Technology, Guangzhou, China"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Weiqiang","family":"Pan","sequence":"additional","affiliation":[{"name":"Information and Network Engineering and Research Centre, South China University of Technology, Guangzhou, China"}],"role":[{"vocabulary":"crossref","role":"author"}]}],"member":"2432","reference":[{"key":"IJDCF.2019040105-0","doi-asserted-by":"publisher","DOI":"10.1109\/BTAS.2013.6712706"},{"key":"IJDCF.2019040105-1","doi-asserted-by":"crossref","unstructured":"Wu, Z., De Leon, P. L., Demiroglu, C., Khodabakhsh, A., King, S., Ling, Z. H., ... & Yamagishi, J. (2016). Anti-spoofing for text-independent speaker verification: An initial database, comparison of countermeasures, and human performance. IEEE\/ACM Transactions on Audio, Speech and Language Processing, 20(8), 768\u2013783.","DOI":"10.1109\/TASLP.2016.2526653"},{"key":"IJDCF.2019040105-2","doi-asserted-by":"publisher","DOI":"10.1121\/1.400476"},{"issue":"2","key":"IJDCF.2019040105-3","first-page":"114","article-title":"Improved closed set text-independent speaker identification by combining MFCC with evidence from flipped filter banks.","volume":"4","author":"S.Chakroborty","year":"2007","journal-title":"International Journal of Signal Processing"},{"key":"IJDCF.2019040105-4","doi-asserted-by":"publisher","DOI":"10.1109\/TASSP.1981.1163530"},{"key":"IJDCF.2019040105-5","doi-asserted-by":"publisher","DOI":"10.1109\/ICASSP.2013.6638975"},{"key":"IJDCF.2019040105-6","article-title":"Perceptual linear predictive (plp) analysis of speech.","author":"H.Hermansky","year":"1990"},{"key":"IJDCF.2019040105-7","first-page":"375","article-title":"Constant-q signal analysis and synthesis.","year":"1978","journal-title":"IEEE International Conference on Acoustics, Speech, and Signal Processing (ICASSP)"},{"key":"IJDCF.2019040105-8","doi-asserted-by":"publisher","DOI":"10.1016\/j.specom.2009.08.009"},{"key":"IJDCF.2019040105-9","unstructured":"Kua, J., Thiruvaran, T., Nosratighods, M. E., Ambikairajah, E., & Epps, J. (2010). Investigation of spectral centroid magnitude and frequency for speaker recognition. In The speaker and language recognition workshop (ODYSSEY)."},{"key":"IJDCF.2019040105-10","doi-asserted-by":"publisher","DOI":"10.1109\/ICASSP.2016.7472724"},{"key":"IJDCF.2019040105-11","first-page":"2062","article-title":"Combining evidences from mel cepstral, cochlear filter cepstral and instantaneous frequency features for detection of natural vs. spoofed speech.","author":"T.Patel","year":"2015","journal-title":"15th Annual Conference of the International Speech Communication Association (INTERSPEECH)"},{"key":"IJDCF.2019040105-12","article-title":"Combining evidences from mel cepstral, cochlear filter cepstral and instantaneous frequency features for detection of natural vs. spoofed speech.","author":"T.Patel","year":"2015","journal-title":"16th Annual Conference of the International Speech Communication Association (INTERSPEECH)"},{"key":"IJDCF.2019040105-13","first-page":"5105","article-title":"Effective of fundament frequency (f0) and strength of excition for spoofed speech detection","author":"T.Patel","year":"2016","journal-title":"IEEE International Conference on Acoustics, Speech, and Signal Processing (ICASSP)"},{"key":"IJDCF.2019040105-14","doi-asserted-by":"publisher","DOI":"10.1109\/JSTSP.2017.2684705"},{"key":"IJDCF.2019040105-15","first-page":"2087","article-title":"A comparison features for synthetic speech detection.","author":"M.Sahidullah","year":"2015","journal-title":"16th Annual Conference of the International Speech Communication Association (INTERSPEECH)"},{"key":"IJDCF.2019040105-16","doi-asserted-by":"publisher","DOI":"10.1109\/ICASSP.1997.596192"},{"key":"IJDCF.2019040105-17","doi-asserted-by":"publisher","DOI":"10.1109\/ASRU.2011.6163899"},{"key":"IJDCF.2019040105-18","first-page":"632","article-title":"Front-end for anti-spoofing countermeasures in speaker verification: Scattering spectral decomposition.","volume":"11","author":"K.Sriskandaraja","year":"2017","journal-title":"IEEE Journal of Selected Topics in Signal Processing"},{"key":"IJDCF.2019040105-19","doi-asserted-by":"crossref","DOI":"10.21437\/Odyssey.2016-41","article-title":"A new feature for automatic speaker verification anti-spoofing: constant Q cepstral coefficients","author":"M.Todisco","year":"2016","journal-title":"The speaker and language recognition workshop"},{"key":"IJDCF.2019040105-20","doi-asserted-by":"crossref","unstructured":"Wu Z. & Li H. (2014). Voice conversation versus speaker verification: an overview. APSIPA Transactions on Signal and Information Processing, 1-16.","DOI":"10.1017\/ATSIP.2014.17"},{"key":"IJDCF.2019040105-21","first-page":"2052","article-title":"Spoofing speech detection using high dimensional magnitude and phase features: The NTU approach for ASVspoof 2015 challenge.","author":"X.Xiao","year":"2015","journal-title":"16th Annual Conference of the International Speech Communication Association (INTERSPEECH)"},{"key":"IJDCF.2019040105-22","doi-asserted-by":"publisher","DOI":"10.1109\/ICASSP.2016.7472636"},{"key":"IJDCF.2019040105-23","doi-asserted-by":"publisher","DOI":"10.1109\/JSTSP.2016.2647199"}],"container-title":["International Journal of Digital Crime and Forensics"],"original-title":[],"language":"ng","link":[{"URL":"https:\/\/www.igi-global.com\/viewtitle.aspx?TitleId=223942","content-type":"unspecified","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2022,5,6]],"date-time":"2022-05-06T21:49:07Z","timestamp":1651873747000},"score":1,"resource":{"primary":{"URL":"https:\/\/services.igi-global.com\/resolvedoi\/resolve.aspx?doi=10.4018\/IJDCF.2019040105"}},"subtitle":[""],"short-title":[],"issued":{"date-parts":[[2019,4,1]]},"references-count":24,"journal-issue":{"issue":"2","published-print":{"date-parts":[[2019,4]]}},"URL":"https:\/\/doi.org\/10.4018\/ijdcf.2019040105","relation":{},"ISSN":["1941-6210","1941-6229"],"issn-type":[{"value":"1941-6210","type":"print"},{"value":"1941-6229","type":"electronic"}],"subject":[],"published":{"date-parts":[[2019,4,1]]}}}