{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2026,4,26]],"date-time":"2026-04-26T03:51:28Z","timestamp":1777175488295,"version":"3.51.4"},"reference-count":27,"publisher":"Springer Science and Business Media LLC","issue":"1","license":[{"start":{"date-parts":[[2023,5,5]],"date-time":"2023-05-05T00:00:00Z","timestamp":1683244800000},"content-version":"tdm","delay-in-days":0,"URL":"https:\/\/creativecommons.org\/licenses\/by\/4.0"},{"start":{"date-parts":[[2023,5,5]],"date-time":"2023-05-05T00:00:00Z","timestamp":1683244800000},"content-version":"vor","delay-in-days":0,"URL":"https:\/\/creativecommons.org\/licenses\/by\/4.0"}],"funder":[{"name":"ANID - Millennium Science Initiative Program","award":["ICN17_002"],"award-info":[{"award-number":["ICN17_002"]}]},{"name":"National Center for Artificial Intelligence CENIA","award":["FB210017"],"award-info":[{"award-number":["FB210017"]}]},{"name":"National Center for Artificial Intelligence CENIA","award":["FB210017"],"award-info":[{"award-number":["FB210017"]}]}],"content-domain":{"domain":["link.springer.com"],"crossmark-restriction":false},"short-container-title":["J AUDIO SPEECH MUSIC PROC."],"abstract":"<jats:title>Abstract<\/jats:title><jats:p>Music inpainting is a sub-task of automated music generation that aims to infill incomplete musical pieces to help musicians in their musical composition process. Many methods have been developed for this task. However, we observe a tendency for each method to be evaluated using different datasets and metrics in the papers where they are presented. This lack of standardization hinders an adequate comparison of these approaches. To tackle these problems, we present MUSIB, a new benchmark for musical score inpainting with standardized conditions for evaluation and reproducibility. MUSIB evaluates four models: Variable Length Piano Infilling (VLI), Music InpaintNet, Music SketchNet, and AnticipationRNN, and over two commonly used datasets: JSB Chorales and IrishFolkSong. We also compile, extend, and propose metrics to adequately quantify note attributes such as pitch and rhythm with <jats:italic>Note Metrics<\/jats:italic>, but also higher-level musical properties with the introduction of <jats:italic>Divergence Metrics<\/jats:italic>, which operate by comparing the distance between distributions of musical features. Our evaluation shows that VLI, a model based on Transformer architecture, is the best performer on a larger dataset, while VAE-based models surpass this Transformer-based model on a relatively small dataset. With MUSIB, we aim at inspiring the community towards better reproducibility in music generation research, setting an example for strongly founded comparisons among SOTA methods.<\/jats:p>","DOI":"10.1186\/s13636-023-00279-6","type":"journal-article","created":{"date-parts":[[2023,5,5]],"date-time":"2023-05-05T09:02:47Z","timestamp":1683277367000},"update-policy":"https:\/\/doi.org\/10.1007\/springer_crossmark_policy","source":"Crossref","is-referenced-by-count":3,"title":["MUSIB: musical score inpainting benchmark"],"prefix":"10.1186","volume":"2023","author":[{"given":"Mauricio","family":"Araneda-Hernandez","sequence":"first","affiliation":[],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Felipe","family":"Bravo-Marquez","sequence":"additional","affiliation":[],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Denis","family":"Parra","sequence":"additional","affiliation":[],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Rodrigo F.","family":"C\u00e1diz","sequence":"additional","affiliation":[],"role":[{"role":"author","vocabulary":"crossref"}]}],"member":"297","published-online":{"date-parts":[[2023,5,5]]},"reference":[{"key":"279_CR1","doi-asserted-by":"publisher","DOI":"10.1007\/978-3-319-70163-9","volume-title":"Deep Learning Techniques for Music Generation","author":"J-P Briot","year":"2020","unstructured":"J.-P. Briot, G. Hadjeres, F.-D. Pachet, Deep Learning Techniques for Music Generation, vol. 1 (Springer, Cham, 2020)"},{"key":"279_CR2","unstructured":"A. Pati, A. Lerch, G. Hadjeres, in Proceedings of the 20th International Society for Music Information Retrieval Conference (ISMIR), Delft, The Netherlands, November 4-8. Learning to Traverse Latent Spaces for Musical Score Inpainting. pp. 343\u2013351"},{"key":"279_CR3","unstructured":"M. F. Dacrema, P. Cremonesi, D. Jannach, in Proceedings of the 13th ACM Conference on Recommender Systems, RecSys 2019, Copenhagen, Denmark, September 16-20, 2019. Are we really making much progress? A worrying analysis of recent neural recommendation approaches, pp. 101\u2013109"},{"key":"279_CR4","doi-asserted-by":"crossref","unstructured":"A. Arango, J. P\u00e9rez, and B. Poblete, in Proceedings of the 42nd International ACM SIGIR Conference on Research and Development in Information Retrieval, SIGIR 2019, Paris, France, July 21-25, 2019. Hate Speech Detection is Not as Easy as You May Think: A Closer Look at Model Validation, pp. 45\u201354","DOI":"10.1145\/3331184.3331262"},{"key":"279_CR5","unstructured":"P. Dhariwal, H. Jun, C. Payne, J.W. Kim, A. Radford, I. Sutskever, Jukebox: A generative model for music. (2020). arXiv preprint arXiv:2005.00341"},{"key":"279_CR6","doi-asserted-by":"crossref","unstructured":"Y. Ren, X. Tan, T. Qin, J. Luan, Z. Zhao, and T.-Y. Liu, in KDD \u201920: The 26th ACM SIGKDD Conference on Knowledge Discovery and Data Mining, Virtual Event, CA, USA, August 23-27, 2020.  DeepSinger: Singing Voice Synthesis with Data Mined From the Web, pp. 1979\u20131989","DOI":"10.1145\/3394486.3403249"},{"key":"279_CR7","unstructured":"Z. Borsos, R. Marinier, D. Vincent, E. Kharitonov, O. Pietquin, M. Sharifi, O. Teboul, D. Grangier, M. Tagliasacchi, N. Zeghidour, Audiolm: a language modeling approach to audio generation. (2022). arXiv preprint arXiv:2209.03143"},{"issue":"1","key":"279_CR8","doi-asserted-by":"publisher","first-page":"89","DOI":"10.5937\/newso1901089B","volume":"53","author":"B Bogunovi\u0107","year":"2019","unstructured":"B. Bogunovi\u0107, Creative cognition in composing music. New Sound Int. J. Music. 53(1), 89\u2013117 (2019)","journal-title":"New Sound Int. J. Music."},{"key":"279_CR9","unstructured":"C.-J. Chang, C.-Y. Lee, and Y.-H. Yang, in Proceedings of the 22nd International Society for Music Information Retrieval Conference, ISMIR 2021, Online, November 7-12, 2021. Variable-Length Music Score Infilling via XLNet and Musically Specialized Positional Encoding, pp. 97\u2013104"},{"key":"279_CR10","unstructured":"K. Chen, C.-I. Wang, T. Berg-Kirkpatrick, and S. Dubnov, in Proceedings of the 21th International Society for Music Information Retrieval Conference, ISMIR 2020, Montreal, Canada, October 11-16, 2020. Music SketchNet: Controllable Music Generation via Factorized Representations of Pitch and Rhythm, pp. 77\u201384"},{"issue":"4","key":"279_CR11","doi-asserted-by":"publisher","first-page":"995","DOI":"10.1007\/s00521-018-3868-4","volume":"32","author":"G Hadjeres","year":"2020","unstructured":"G. Hadjeres, F. Nielsen, Anticipation-rnn: Enforcing unary constraints in sequence generation, with application to interactive music generation. Neural Computing and Applications. 32(4), 995\u20131005 (2020)","journal-title":"Neural Computing and Applications."},{"key":"279_CR12","unstructured":"M.S. Cuthbert, C. Ariza, in Proceedings of the 11th International Society for Music Information Retrieval Conference, ISMIR 2010, Utrecht, August 9-13, 2010, ed. by J.S. Downie, R.C. Veltkamp. Music21: A toolkit for computer-aided musicology and symbolic music data. (International Society for Music Information Retrieval, 2010), pp. 637\u2013642. http:\/\/ismir2010.ismir.net\/proceedings\/ismir2010-108.pdf"},{"key":"279_CR13","unstructured":"B.L. Sturm, J.F. Santos, O. Ben-Tal, I. Korshunova, Music transcription modelling and composition using deep learning. CoRR abs\/1604.08723 (2016). http:\/\/arxiv.org\/abs\/1604.08723"},{"key":"279_CR14","unstructured":"G. Hadjeres, F. Pachet, F. Nielsen, in International Conference on Machine Learning. Sydney. Deepbach: a steerable model for bach chorales generation. (PMLR, 2017), pp. 1362\u20131371"},{"key":"279_CR15","unstructured":"C.-Z. A. Huang, T. Cooijmans, A. Roberts, A. C. Courville, and D. Eck,  in Proceedings of the 18th International Society for Music Information Retrieval Conference, ISMIR 2017, Suzhou, China, October 23-27, 2017. Counterpoint by Convolution, pp. 211\u2013218"},{"key":"279_CR16","unstructured":"D.P. Kingma, M. Welling, in 2nd International Conference on Learning Representations, ICLR 2014, Banff, AB, Canada, April 14-16, 2014, Conference Track Proceedings, ed. by Y. Bengio, Y. LeCun. Auto-encoding variational bayes. (2014)"},{"key":"279_CR17","unstructured":"Z. Yang, Z. Dai, Y. Yang, J. Carbonell, R.R. Salakhutdinov, Q.V. Le, Xlnet: Generalized autoregressive pretraining for language understanding. Adv. Neural Inf. Process. Syst. 32,\u00a05754\u20135764\u00a0(2019)"},{"key":"279_CR18","unstructured":"S. Wu, Y. Yang, in Proceedings of the 21th International Society for Music Information Retrieval Conference, ISMIR 2020, Montreal, Canada, October 11-16, 2020. ed. by J. Cumming, J.H. Lee, B. McFee, M. Schedl, J. Devaney, C. McKay, E. Zangerle, T. de\u00a0Reuse. The jazz transformer on the front line: Exploring the shortcomings of ai-composed music through quantitative measures. (2020) pp. 142\u2013149"},{"key":"279_CR19","unstructured":"M. Cartagena, Exploring symbolic music generation techniques using conditional generative adversarial networks. Master\u2019s thesis, Escuela de Ingenier\u00eda. Pontificia Universidad Cat\u00f3lica de chile. (2021) https:\/\/repositorio.uc.cl\/handle\/11534\/60993"},{"issue":"1","key":"279_CR20","doi-asserted-by":"publisher","first-page":"145","DOI":"10.1109\/18.61115","volume":"37","author":"J Lin","year":"1991","unstructured":"J. Lin, Divergence measures based on the shannon entropy. IEEE Trans. Inf. Theory. 37(1), 145\u2013151 (1991). https:\/\/doi.org\/10.1109\/18.61115","journal-title":"IEEE Trans. Inf. Theory."},{"key":"279_CR21","doi-asserted-by":"publisher","first-page":"1","DOI":"10.1007\/978-3-642-00234-2","volume-title":"Encyclopedia of Distances","author":"MM Deza","year":"2009","unstructured":"M.M. Deza, E. Deza, Encyclopedia of distances, in Encyclopedia of Distances. (Springer, Cham, 2009), pp.1\u2013583"},{"issue":"2","key":"279_CR22","doi-asserted-by":"publisher","first-page":"221","DOI":"10.3390\/e22020221","volume":"22","author":"F Nielsen","year":"2020","unstructured":"F. Nielsen, On a generalization of the jensen-shannon divergence and the jensen-shannon centroid. Entropy. 22(2), 221 (2020)","journal-title":"Entropy."},{"key":"279_CR23","unstructured":"A. Dosovitskiy, L. Beyer, A. Kolesnikov, D. Weissenborn, X. Zhai, T. Unterthiner, M. Dehghani, M. Minderer, G. Heigold, S. Gelly, J. Uszkoreit, N. Houlsby, in 9th International Conference on Learning Representations, ICLR 2021, Virtual Event, Austria, May 3-7, 2021. An image is worth 16x16 words: Transformers for image recognition at scale. (OpenReview.net, Austria, 2021). https:\/\/openreview.net\/forum?id=YicbFdNTTy"},{"key":"279_CR24","first-page":"6840","volume":"33","author":"J Ho","year":"2020","unstructured":"J. Ho, A. Jain, P. Abbeel, Denoising diffusion probabilistic models. Adv. Neural Inf. Process. Syst. 33, 6840\u20136851 (2020)","journal-title":"Adv. Neural Inf. Process. Syst."},{"key":"279_CR25","doi-asserted-by":"crossref","unstructured":"K. Kilgour, M. Zuluaga, D. Roblek, and M. Sharifi, in Interspeech 2019, 20th Annual Conference of the International Speech Communication Association, Graz, Austria, 15-19 September 2019. Fr\u00e9chet Audio Distance: A Reference-Free Metric for Evaluating Music Enhancement Algorithms, pp. 2350\u20132354","DOI":"10.21437\/Interspeech.2019-2219"},{"key":"279_CR26","unstructured":"F. T. Liang, M. Gotham, M. Johnson, and J. Shotton, in Proceedings of the 18th International Society for Music Information Retrieval Conference, ISMIR 2017, Suzhou, China, October 23-27, 2017. Automatic Stylistic Composition of Bach Chorales with Deep LSTM, pp. 449\u2013456"},{"issue":"9","key":"279_CR27","doi-asserted-by":"publisher","first-page":"4773","DOI":"10.1007\/s00521-018-3849-7","volume":"32","author":"L-C Yang","year":"2020","unstructured":"L.-C. Yang, A. Lerch, On the evaluation of generative models in music. Neural Comput. Appl. 32(9), 4773\u20134784 (2020)","journal-title":"Neural Comput. Appl."}],"container-title":["EURASIP Journal on Audio, Speech, and Music Processing"],"original-title":[],"language":"en","link":[{"URL":"https:\/\/link.springer.com\/content\/pdf\/10.1186\/s13636-023-00279-6.pdf","content-type":"application\/pdf","content-version":"vor","intended-application":"text-mining"},{"URL":"https:\/\/link.springer.com\/article\/10.1186\/s13636-023-00279-6\/fulltext.html","content-type":"text\/html","content-version":"vor","intended-application":"text-mining"},{"URL":"https:\/\/link.springer.com\/content\/pdf\/10.1186\/s13636-023-00279-6.pdf","content-type":"application\/pdf","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2023,5,5]],"date-time":"2023-05-05T09:06:37Z","timestamp":1683277597000},"score":1,"resource":{"primary":{"URL":"https:\/\/asmp-eurasipjournals.springeropen.com\/articles\/10.1186\/s13636-023-00279-6"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2023,5,5]]},"references-count":27,"journal-issue":{"issue":"1","published-online":{"date-parts":[[2023,12]]}},"alternative-id":["279"],"URL":"https:\/\/doi.org\/10.1186\/s13636-023-00279-6","relation":{},"ISSN":["1687-4722"],"issn-type":[{"value":"1687-4722","type":"electronic"}],"subject":[],"published":{"date-parts":[[2023,5,5]]},"assertion":[{"value":"13 September 2022","order":1,"name":"received","label":"Received","group":{"name":"ArticleHistory","label":"Article History"}},{"value":"14 March 2023","order":2,"name":"accepted","label":"Accepted","group":{"name":"ArticleHistory","label":"Article History"}},{"value":"5 May 2023","order":3,"name":"first_online","label":"First Online","group":{"name":"ArticleHistory","label":"Article History"}},{"order":1,"name":"Ethics","group":{"name":"EthicsHeading","label":"Declarations"}},{"value":"This article does not contain any studies with human participants or animals performed by any of the authors.","order":2,"name":"Ethics","group":{"name":"EthicsHeading","label":"Ethics approval and consent to participate"}},{"value":"Not applicable.","order":3,"name":"Ethics","group":{"name":"EthicsHeading","label":"Consent for publication"}},{"value":"The authors declare that they have no competing interests.","order":4,"name":"Ethics","group":{"name":"EthicsHeading","label":"Competing interests"}}],"article-number":"19"}}