{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2026,1,29]],"date-time":"2026-01-29T22:55:40Z","timestamp":1769727340225,"version":"3.49.0"},"reference-count":26,"publisher":"Institute of Electronics, Information and Communications Engineers (IEICE)","issue":"3","content-domain":{"domain":[],"crossmark-restriction":false},"short-container-title":["IEICE Trans. Inf. &amp; Syst."],"published-print":{"date-parts":[[2020,3,1]]},"DOI":"10.1587\/transinf.2019edp7175","type":"journal-article","created":{"date-parts":[[2020,2,29]],"date-time":"2020-02-29T22:10:51Z","timestamp":1583014251000},"page":"695-701","source":"Crossref","is-referenced-by-count":23,"title":["Combining CNN and Broad Learning for Music Classification"],"prefix":"10.1587","volume":"E103.D","author":[{"given":"Huan","family":"TANG","sequence":"first","affiliation":[{"name":"School of Information Science and Engineering, East China University of Science and Technology"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Ning","family":"CHEN","sequence":"additional","affiliation":[{"name":"School of Information Science and Engineering, East China University of Science and Technology"}],"role":[{"role":"author","vocabulary":"crossref"}]}],"member":"532","reference":[{"key":"1","unstructured":"[1] K. Choi, G. Fazekas, and M. Sandler, \u201cAutomatic tagging using deep convolutional neural networks,\u201d 17th International Society of Music Information Retrieval (ISMIR), pp.805-811, Aug. 2016."},{"key":"2","doi-asserted-by":"crossref","unstructured":"[2] K. Choi, G. Fazekas, M. Sandler, and K. Cho, \u201cConvolutional recurrent neural networks for music classification,\u201d 16th IEEE International Conference on Acoustics, Speech and Signal Processing (ICASSP), pp.2392-2396, March 2017. 10.1109\/icassp.2017.7952585","DOI":"10.1109\/ICASSP.2017.7952585"},{"key":"3","doi-asserted-by":"publisher","unstructured":"[3] E.J. Humphrey, J.P. Bello, and Y. LeCun, \u201cFeature learning and deep architectures: new directions for music informatics,\u201d Journal of Intelligent Information Systems, vol.41, no.3, pp.461-481, 2013. 10.1007\/s10844-013-0248-5","DOI":"10.1007\/s10844-013-0248-5"},{"key":"4","unstructured":"[4] N. Chen and S. Wang, \u201cHigh-level music descriptor extraction algorithm based on combination of multi-channel CNNs and LSTM,\u201d 18th International Society of Music Information Retrieval (ISMIR), pp.509-514, Oct. 2017."},{"key":"5","unstructured":"[5] J. Pons, O. Nieto, M. Prockup, E.M. Schmidt, A.F. Ehmann, and X. Serra, \u201cEnd-to-end learning for music audio tagging at scale,\u201d 19th International Society for Music Information Retrieval Conference (ISMIR), pp.637-644, Sept. 2018."},{"key":"6","doi-asserted-by":"crossref","unstructured":"[6] J. Schluter and S. Bock, \u201cImproved musical onset detection with convolutional neural networks,\u201d IEEE International Conference on Acoustics, Speech and Signal Processing (ICASSP), pp.6979-6983, May 2014. 10.1109\/icassp.2014.6854953","DOI":"10.1109\/ICASSP.2014.6854953"},{"key":"7","doi-asserted-by":"crossref","unstructured":"[7] M. Malik, S. Adavanne, K. Drossos, T. Virtanen, D. Ticha, and R. Jarina, \u201cStacked convolutional and recurrent neural networks for music emotion recognition,\u201d Sound and Music Computing (SMC) Conference, pp.208-214, Oct. 2017.","DOI":"10.23919\/EUSIPCO.2017.8081505"},{"key":"8","doi-asserted-by":"publisher","unstructured":"[8] J. Deng and Y.-K. Kwok, \u201cLarge vocabulary automatic chord estimation using bidirectional long short-term memory recurrent neural network with even chance training,\u201d Journal of New Music Research, vol.47, no.1, pp.53-67, 2017. 10.1080\/09298215.2017.1367820","DOI":"10.1080\/09298215.2017.1367820"},{"key":"9","unstructured":"[9] S. Stober, D.J. Cameron, and J.A. Grahn, \u201cUsing convolutional neural networks to recognize rhythm stimuli from electroencephalography recordings,\u201d Advances in Neural Information Processing Systems (NIPS), pp.1449-1457, Dec. 2014."},{"key":"10","unstructured":"[10] D. Stoller, S. Ewert, and S. Dixon, \u201cWave-u-net: a multi-scale neural network for end-to-end audio source separation,\u201d 19th International Society for Music Information Retrieval Conference (ISMIR), pp.334-340, Sept. 2018."},{"key":"11","doi-asserted-by":"crossref","unstructured":"[11] A. Abdul, J. Chen, H.Y. Liao, and S.H. Chang, \u201cAn emotion-aware personalized music recommendation system using a convolutional neural networks approach,\u201d Applied Sciences, vol.8, no.7, pp.1-16, 2018.","DOI":"10.3390\/app8071103"},{"key":"12","doi-asserted-by":"crossref","unstructured":"[12] J. Dai, S. Liang, W. Xue, C.J. Ni, and W.J. Liu, \u201cLong short-term memory recurrent neural network based segment features for music genre classification,\u201d International Symposium on Chinese Spoken Language Processing (ISCSLP), pp.1-5, IEEE, 2016. 10.1109\/iscslp.2016.7918369","DOI":"10.1109\/ISCSLP.2016.7918369"},{"key":"13","unstructured":"[13] L. Feng, S. Liu, and J. Yao, \u201cMusic genre classification with paralleling recurrent convolutional neural network,\u201d arXiv preprint arXiv:1712.08370, 2017."},{"key":"14","doi-asserted-by":"crossref","unstructured":"[14] R.L. Aguiar, Y.M.G. Costa, and C.N. Silla, \u201cExploring data augmentation to improve music genre classification with ConvNets,\u201d 2018 International Joint Conference on Neural Networks (IJCNN), pp.1-8, IEEE, July 2018. 10.1109\/ijcnn.2018.8489166","DOI":"10.1109\/IJCNN.2018.8489166"},{"key":"15","doi-asserted-by":"crossref","unstructured":"[15] T. Raissi, A. Tibo, and P. Bientinesi, \u201cExtended pipeline for content-based feature engineering in music genre recognition,\u201d 43rd International Conference on Acoustics, Speech, and Signal Processing (ICASSP), pp.2661-2665, IEEE, April 2018. 10.1109\/icassp.2018.8461807","DOI":"10.1109\/ICASSP.2018.8461807"},{"key":"16","doi-asserted-by":"crossref","unstructured":"[16] J. Pons and X. Serra, \u201cRandomly weighted CNNs for (music) audio classification,\u201d 44th International Conference on Acoustics, Speech and Signal Processing (ICASSP), pp.336-340, IEEE, May 2019. 10.1109\/icassp.2019.8682912","DOI":"10.1109\/ICASSP.2019.8682912"},{"key":"17","doi-asserted-by":"crossref","unstructured":"[17] J. Pons, T. Lidy, and X. Serra, \u201cExperimenting with musically motivated convolutional neural networks,\u201d 14th IEEE International Workshop on Content-Based Multimedia Indexing (CBMI), pp.1-6, June 2016. 10.1109\/cbmi.2016.7500246","DOI":"10.1109\/CBMI.2016.7500246"},{"key":"18","doi-asserted-by":"crossref","unstructured":"[18] K.H. Wong, C.P. Tang, K.L. Chui, Y.K. Yu, Z.L. Zeng, X. Jiang, G. Chen, and Z. Chen, \u201cMusic genre classification using a hierarchical long short term memory (LSTM) model,\u201d 3rd International Workshop on Pattern Recognition (IWPR), vol.10828, pp.108281B, May 2018. 10.1117\/12.2501763","DOI":"10.1117\/12.2501763"},{"key":"19","unstructured":"[19] K. Kowsari, M. Heidarysafa, D.E. Brown, K.J. Meimandi, and L.E. Barnes, \u201cRMDL: random multimodel deep learning for classification,\u201d International Conference on Information System and Data Mining (ICISDM), pp.19-28, April 2018. 10.1145\/3206098.3206111"},{"key":"20","doi-asserted-by":"publisher","unstructured":"[20] C.L.P. Chen and Z.L. Liu, \u201cBroad learning system: an effective and efficient incremental learning system without the need for deep architecture,\u201d IEEE Trans. Neural Netw. Learn. Syst., vol.29, no.1, pp.10-24, 2018. 10.1109\/tnnls.2017.2716952","DOI":"10.1109\/TNNLS.2017.2716952"},{"key":"21","unstructured":"[21] N. Srivastava, G. Hinton, A. Krizhevsky, I. Sutskever, and R. Salakhutdinov, \u201cDropout: a simple way to prevent neural networks from overfitting,\u201d Journal of Machine Learning Research, vol.15, no.1, pp.1929-1958, 2014."},{"key":"22","unstructured":"[22] K. Simonyan and A. Zisserman, \u201cVery deep convolutional networks for large-scale image recognition,\u201d 3rd International Conference on Learning Representations (ICLR), May 2015."},{"key":"23","unstructured":"[23] D.P. Kingma and J. Ba, \u201cAdam: A method for stochastic optimization,\u201d The 3rd International Conference for Learning Representations (ICLR), pp.1-15, May 2015."},{"key":"24","doi-asserted-by":"crossref","unstructured":"[24] G. Tzanetakis and P. Cook, \u201cMusical genre classification of audio signals,\u201d IEEE Transactions on Speech and Audio Processing, vol.10, no.5, pp.293-302, 2002. 10.1109\/tsa.2002.800560","DOI":"10.1109\/TSA.2002.800560"},{"key":"25","unstructured":"[25] F. Gouyon, S. Dixon, E. Pampalk, and G. Widmer, \u201cEvaluating rhythmic descriptors for musical genre classification,\u201d 25th International Conference on Audio Engineering Society (AES), pp.196-204, 2004."},{"key":"26","doi-asserted-by":"crossref","unstructured":"[26] V. Kirandziska and N. Ackovska, \u201cFinding important sound features for emotion evaluation classification,\u201d European Conference on Electronics (Eurocon), IEEE, pp.1637-1644, July 2013. 10.1109\/eurocon.2013.6625196","DOI":"10.1109\/EUROCON.2013.6625196"}],"container-title":["IEICE Transactions on Information and Systems"],"original-title":[],"language":"en","link":[{"URL":"https:\/\/www.jstage.jst.go.jp\/article\/transinf\/E103.D\/3\/E103.D_2019EDP7175\/_pdf","content-type":"unspecified","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2020,3,7]],"date-time":"2020-03-07T03:26:48Z","timestamp":1583551608000},"score":1,"resource":{"primary":{"URL":"https:\/\/www.jstage.jst.go.jp\/article\/transinf\/E103.D\/3\/E103.D_2019EDP7175\/_article"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2020,3,1]]},"references-count":26,"journal-issue":{"issue":"3","published-print":{"date-parts":[[2020]]}},"URL":"https:\/\/doi.org\/10.1587\/transinf.2019edp7175","relation":{},"ISSN":["0916-8532","1745-1361"],"issn-type":[{"value":"0916-8532","type":"print"},{"value":"1745-1361","type":"electronic"}],"subject":[],"published":{"date-parts":[[2020,3,1]]}}}