{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2025,6,19]],"date-time":"2025-06-19T04:10:44Z","timestamp":1750306244631,"version":"3.41.0"},"reference-count":29,"publisher":"Association for Computing Machinery (ACM)","issue":"1","license":[{"start":{"date-parts":[[2016,12,16]],"date-time":"2016-12-16T00:00:00Z","timestamp":1481846400000},"content-version":"vor","delay-in-days":0,"URL":"https:\/\/www.acm.org\/publications\/policies\/copyright_policy#Background"}],"funder":[{"name":"CAPES Foundation"},{"name":"Ministry of Education of Brazil","award":["BEX-2106\/13-2"],"award-info":[{"award-number":["BEX-2106\/13-2"]}]}],"content-domain":{"domain":["dl.acm.org"],"crossmark-restriction":true},"short-container-title":["ACM Trans. Multimedia Comput. Commun. Appl."],"published-print":{"date-parts":[[2017,2,28]]},"abstract":"<jats:p>Solf\u00e8ge is a general technique used in the music learning process that involves the vocal performance of melodies, regarding the time and duration of musical sounds as specified in the music score, properly associated with the meter-mimicking performed by hand movement. This article presents an audiovisual approach for automatic assessment of this relevant musical study practice. The proposed system combines the gesture of meter-mimicking (video information) with the melodic transcription (audio information), where hand movement works as a metronome, controlling the time flow (tempo) of the musical piece. Thus, meter-mimicking is used to align the music score (ground truth) with the sung melody, allowing assessment even in time-dynamic scenarios. Audio analysis is applied to achieve the melodic transcription of the sung notes and the solf\u00e8ge performances are evaluated by a set of Bayesian classifiers that were generated from real evaluations done by experts listeners.<\/jats:p>","DOI":"10.1145\/3007194","type":"journal-article","created":{"date-parts":[[2016,12,19]],"date-time":"2016-12-19T17:06:31Z","timestamp":1482167191000},"page":"1-21","update-policy":"https:\/\/doi.org\/10.1145\/crossmark-policy","source":"Crossref","is-referenced-by-count":1,"title":["Audiovisual Tool for Solf\u00e8ge Assessment"],"prefix":"10.1145","volume":"13","author":[{"given":"Rodrigo","family":"Schramm","sequence":"first","affiliation":[{"name":"Universidade Federal do Rio Grande do Sul, RS, Brazil"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Helena De Souza","family":"Nunes","sequence":"additional","affiliation":[{"name":"Universidade Federal da Bahia, BA, Brazil"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Cl\u00e1udio Rosito","family":"Jung","sequence":"additional","affiliation":[{"name":"Universidade Federal do Rio Grande do Sul, RS, Brazil"}],"role":[{"role":"author","vocabulary":"crossref"}]}],"member":"320","published-online":{"date-parts":[[2016,12,16]]},"reference":[{"key":"e_1_2_1_1_1","doi-asserted-by":"publisher","DOI":"10.1007\/978-3-642-12553-9_7"},{"key":"e_1_2_1_2_1","first-page":"4","article-title":"YIN, a fundamental frequency estimator for speech and music","volume":"111","author":"de Cheveign\u00e9 Alain","year":"2002","unstructured":"Alain de Cheveign\u00e9 and Hideki Kawahara . 2002 . YIN, a fundamental frequency estimator for speech and music . J. Acoust. Soc. Am. 111 , 4 (Apr. 2002), 1917--1930. Alain de Cheveign\u00e9 and Hideki Kawahara. 2002. YIN, a fundamental frequency estimator for speech and music. J. Acoust. Soc. Am. 111, 4 (Apr. 2002), 1917--1930.","journal-title":"J. Acoust. Soc. Am."},{"key":"e_1_2_1_3_1","volume-title":"Stork","author":"Duda Richard O.","year":"2001","unstructured":"Richard O. Duda , Peter E. Hart , and David G . Stork . 2001 . Pattern Classification (2nd ed.). Wiley-Interscience . Richard O. Duda, Peter E. Hart, and David G. Stork. 2001. Pattern Classification (2nd ed.). Wiley-Interscience."},{"key":"e_1_2_1_4_1","doi-asserted-by":"publisher","DOI":"10.1162\/COMJ_a_00180"},{"key":"e_1_2_1_5_1","unstructured":"Emile Jaques-Dalcroze. 2014. Rhythm Music and Education. Read Books Ltd.  Emile Jaques-Dalcroze. 2014. Rhythm Music and Education. Read Books Ltd."},{"volume-title":"Proceedings of First International Conference on Data Mining (SDM\u201901)","author":"Eamonn","key":"e_1_2_1_6_1","unstructured":"Eamonn J. Keogh and Michael J. Pazzani. 2001. Derivative dynamic time warping . In Proceedings of First International Conference on Data Mining (SDM\u201901) . Eamonn J. Keogh and Michael J. Pazzani. 2001. Derivative dynamic time warping. In Proceedings of First International Conference on Data Mining (SDM\u201901)."},{"key":"e_1_2_1_7_1","doi-asserted-by":"publisher","DOI":"10.1214\/aos\/1176348783"},{"volume-title":"Signal Processing Methods for Music Transcription","author":"Klapuri Anssi","key":"e_1_2_1_8_1","unstructured":"Anssi Klapuri and Manuel Davy . 2006. Signal Processing Methods for Music Transcription . Springer-Verlag , New York, NY . Anssi Klapuri and Manuel Davy. 2006. Signal Processing Methods for Music Transcription. Springer-Verlag, New York, NY."},{"key":"e_1_2_1_10_1","doi-asserted-by":"publisher","DOI":"10.1109\/ICOT.2014.6956625"},{"key":"e_1_2_1_11_1","doi-asserted-by":"publisher","DOI":"10.1080\/10447318.2012.720197"},{"key":"e_1_2_1_12_1","volume-title":"Proceedings of the 9th Sound and Music Computing Conference, Sound and Music Computing (Eds.). 271--276","author":"Mandanici Marcella","year":"2012","unstructured":"Marcella Mandanici and Sylviane Sapir . 2012 . Disembodied voices: A kinect virtual choir conductor . In Proceedings of the 9th Sound and Music Computing Conference, Sound and Music Computing (Eds.). 271--276 . Marcella Mandanici and Sylviane Sapir. 2012. Disembodied voices: A kinect virtual choir conductor. In Proceedings of the 9th Sound and Music Computing Conference, Sound and Music Computing (Eds.). 271--276."},{"key":"e_1_2_1_13_1","doi-asserted-by":"publisher","DOI":"10.1109\/ICASSP.2014.6853678"},{"key":"e_1_2_1_14_1","volume-title":"Proceedings of the 15th International Society for Music Information Retrieval Conference (ISMIR\u201914)","author":"Molina Emilio","year":"2014","unstructured":"Emilio Molina , Ana M. Barbancho , Lorenzo J. Tard\u00f3n , and Isabel Barbancho . 2014 . Evaluation framework for automatic singing transcription . In Proceedings of the 15th International Society for Music Information Retrieval Conference (ISMIR\u201914) . ISMIR, 567--572. Emilio Molina, Ana M. Barbancho, Lorenzo J. Tard\u00f3n, and Isabel Barbancho. 2014. Evaluation framework for automatic singing transcription. In Proceedings of the 15th International Society for Music Information Retrieval Conference (ISMIR\u201914). ISMIR, 567--572."},{"volume-title":"Proceedings of IEEE International Conference on Acoustics, Speech and Signal Processing (ICASSP\u201913)","author":"Molina Emilio","key":"e_1_2_1_15_1","unstructured":"Emilio Molina , Isabel Barbancho , Emilia G\u00f3mez , Ana M. Barbancho , and Lorenzo J. Tard\u00f3n . 2013. Fundamental frequency alignment vs. note-based melodic similarity for singing voice assessment . In Proceedings of IEEE International Conference on Acoustics, Speech and Signal Processing (ICASSP\u201913) . 744--748. Emilio Molina, Isabel Barbancho, Emilia G\u00f3mez, Ana M. Barbancho, and Lorenzo J. Tard\u00f3n. 2013. Fundamental frequency alignment vs. note-based melodic similarity for singing voice assessment. In Proceedings of IEEE International Conference on Acoustics, Speech and Signal Processing (ICASSP\u201913). 744--748."},{"key":"e_1_2_1_16_1","doi-asserted-by":"publisher","DOI":"10.1109\/TASLP.2014.2331102"},{"volume-title":"Information Retrieval for Music and Motion","author":"M\u00fcller Meinard","key":"e_1_2_1_17_1","unstructured":"Meinard M\u00fcller . 2007. Information Retrieval for Music and Motion . Springer-Verlag , Berlin . Meinard M\u00fcller. 2007. Information Retrieval for Music and Motion. Springer-Verlag, Berlin."},{"volume-title":"Fundamentals of Music Processing --","author":"M\u00fcller Meinard","key":"e_1_2_1_18_1","unstructured":"Meinard M\u00fcller . 2015. Fundamentals of Music Processing -- . Springer-Verlag , Berlin . Meinard M\u00fcller. 2015. Fundamentals of Music Processing -- . Springer-Verlag, Berlin."},{"volume-title":"The Analysis and Cognition of Basic Melodic Structures: The Implication-Realization Model","author":"Narmour Eugene","key":"e_1_2_1_19_1","unstructured":"Eugene Narmour . 1990. The Analysis and Cognition of Basic Melodic Structures: The Implication-Realization Model . The University of Chicago Press . Eugene Narmour. 1990. The Analysis and Cognition of Basic Melodic Structures: The Implication-Realization Model. The University of Chicago Press."},{"key":"e_1_2_1_20_1","volume-title":"The Grammar of Conducting","author":"Rudolf Max","unstructured":"Max Rudolf . 1980. The Grammar of Conducting ( 2 nd ed.). Schirmer Books Inc ., New York, NY. Max Rudolf. 1980. The Grammar of Conducting (2nd ed.). Schirmer Books Inc., New York, NY.","edition":"2"},{"key":"e_1_2_1_21_1","volume-title":"Proceedings of ISCA\u2014Tutorial and Research Workshop on Statistical and Perceptual Audio. MIT Press.","author":"Ryyn\u00e4nen Matti","year":"2004","unstructured":"Matti Ryyn\u00e4nen and Anssi Klapuri . 2004 . Modelling of note events for singing transcription . In Proceedings of ISCA\u2014Tutorial and Research Workshop on Statistical and Perceptual Audio. MIT Press. Matti Ryyn\u00e4nen and Anssi Klapuri. 2004. Modelling of note events for singing transcription. In Proceedings of ISCA\u2014Tutorial and Research Workshop on Statistical and Perceptual Audio. MIT Press."},{"key":"e_1_2_1_22_1","volume-title":"Proceedings of the 16th International Society for Music Information Retrieval Conference (ISMIR\u201915)","author":"Schramm Rodrigo","year":"2015","unstructured":"Rodrigo Schramm , Helena de Souza Nunes , and Cl\u00e1udio Rosito Jung . 2015 a. Automatic Solf\u00e8ge assessment . In Proceedings of the 16th International Society for Music Information Retrieval Conference (ISMIR\u201915) . 183--189. Rodrigo Schramm, Helena de Souza Nunes, and Cl\u00e1udio Rosito Jung. 2015a. Automatic Solf\u00e8ge assessment. In Proceedings of the 16th International Society for Music Information Retrieval Conference (ISMIR\u201915). 183--189."},{"key":"e_1_2_1_23_1","doi-asserted-by":"publisher","DOI":"10.1109\/TMM.2014.2377553"},{"volume-title":"Analysis and Music Education","author":"Swanwick Keith","key":"e_1_2_1_24_1","unstructured":"Keith Swanwick . 1994. Musical Knowledge , Intuition , Analysis and Music Education . Routledge , Londres . Keith Swanwick. 1994. Musical Knowledge, Intuition, Analysis and Music Education. Routledge, Londres."},{"key":"e_1_2_1_25_1","doi-asserted-by":"publisher","DOI":"10.1214\/aoms\/1177728730"},{"key":"e_1_2_1_26_1","volume-title":"Proceedings of the IEEE International Conference on Multimedia and Expo (ICME). 1--6.","author":"Toh Leng-Wee","year":"2013","unstructured":"Leng-Wee Toh , W. Chao , and Yi-Shin Chen . 2013 . An interactive conducting system using kinect . In Proceedings of the IEEE International Conference on Multimedia and Expo (ICME). 1--6. Leng-Wee Toh, W. Chao, and Yi-Shin Chen. 2013. An interactive conducting system using kinect. In Proceedings of the IEEE International Conference on Multimedia and Expo (ICME). 1--6."},{"key":"e_1_2_1_27_1","volume-title":"Proceedings of the 2003 Finnish Signal Processing Symposium. 59--63","author":"Viitaniemi Timo","year":"2003","unstructured":"Timo Viitaniemi , Anssi Klapuri , and Antti Eronen . 2003 . A probabilistic model for the transcription of single-voice melodies . In Proceedings of the 2003 Finnish Signal Processing Symposium. 59--63 . Timo Viitaniemi, Anssi Klapuri, and Antti Eronen. 2003. A probabilistic model for the transcription of single-voice melodies. In Proceedings of the 2003 Finnish Signal Processing Symposium. 59--63."},{"key":"e_1_2_1_28_1","volume-title":"Statistical Pattern Recognition","author":"Webb Andrew R.","unstructured":"Andrew R. Webb . 2011. Statistical Pattern Recognition ( 3 rd ed.). Wiley, Chichester , UK. Andrew R. Webb. 2011. Statistical Pattern Recognition (3rd ed.). Wiley, Chichester, UK.","edition":"3"},{"volume-title":"Proceedings of American Control Conference. 2864--2869","author":"Zhang Yang","key":"e_1_2_1_29_1","unstructured":"Yang Zhang and T. F. Edgar . 2008. A robust dynamic time warping algorithm for batch trajectory synchronization . In Proceedings of American Control Conference. 2864--2869 . Yang Zhang and T. F. Edgar. 2008. A robust dynamic time warping algorithm for batch trajectory synchronization. In Proceedings of American Control Conference. 2864--2869."},{"volume-title":"Assessment in Music Education: From Policy to Practice, Don Lebler, Gemmal Carey, and Scott D","author":"Zhukov Katie","key":"e_1_2_1_30_1","unstructured":"Katie Zhukov . 2015. Challenging approaches to assessment of instrumental learning . In Assessment in Music Education: From Policy to Practice, Don Lebler, Gemmal Carey, and Scott D . Harrison (Eds.). Vol. 16 . Springer International , Switzerland . Katie Zhukov. 2015. Challenging approaches to assessment of instrumental learning. In Assessment in Music Education: From Policy to Practice, Don Lebler, Gemmal Carey, and Scott D. Harrison (Eds.). Vol. 16. Springer International, Switzerland."}],"container-title":["ACM Transactions on Multimedia Computing, Communications, and Applications"],"original-title":[],"language":"en","link":[{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/3007194","content-type":"unspecified","content-version":"vor","intended-application":"text-mining"},{"URL":"https:\/\/dl.acm.org\/doi\/pdf\/10.1145\/3007194","content-type":"unspecified","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2025,6,18]],"date-time":"2025-06-18T04:23:41Z","timestamp":1750220621000},"score":1,"resource":{"primary":{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/3007194"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2016,12,16]]},"references-count":29,"journal-issue":{"issue":"1","published-print":{"date-parts":[[2017,2,28]]}},"alternative-id":["10.1145\/3007194"],"URL":"https:\/\/doi.org\/10.1145\/3007194","relation":{},"ISSN":["1551-6857","1551-6865"],"issn-type":[{"type":"print","value":"1551-6857"},{"type":"electronic","value":"1551-6865"}],"subject":[],"published":{"date-parts":[[2016,12,16]]},"assertion":[{"value":"2016-03-01","order":0,"name":"received","label":"Received","group":{"name":"publication_history","label":"Publication History"}},{"value":"2016-10-01","order":1,"name":"accepted","label":"Accepted","group":{"name":"publication_history","label":"Publication History"}},{"value":"2016-12-16","order":2,"name":"published","label":"Published","group":{"name":"publication_history","label":"Publication History"}}]}}