{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2026,8,4]],"date-time":"2026-08-04T03:20:03Z","timestamp":1785813603316,"version":"3.56.0"},"publisher-location":"New York, NY, USA","reference-count":37,"publisher":"ACM","license":[{"start":{"date-parts":[[2022,10,10]],"date-time":"2022-10-10T00:00:00Z","timestamp":1665360000000},"content-version":"vor","delay-in-days":0,"URL":"https:\/\/www.acm.org\/publications\/policies\/copyright_policy#Background"}],"funder":[{"name":"Project of Key Laboratory of Intelligent Processing Technology for Digital Music (Zhejiang Conservatory of Music), Ministry of Culture and Tourism","award":["2022DMKLB001"],"award-info":[{"award-number":["2022DMKLB001"]}]},{"name":"Key R&D Program of Zhejiang Province","award":["2022C03126"],"award-info":[{"award-number":["2022C03126"]}]},{"name":"Key Project of Natural Science Foundation of Zhejiang Province","award":["LZ19F020002"],"award-info":[{"award-number":["LZ19F020002"]}]}],"content-domain":{"domain":["dl.acm.org"],"crossmark-restriction":true},"short-container-title":[],"published-print":{"date-parts":[[2022,10,10]]},"DOI":"10.1145\/3503161.3548357","type":"proceedings-article","created":{"date-parts":[[2022,10,10]],"date-time":"2022-10-10T15:43:12Z","timestamp":1665416592000},"page":"1047-1056","update-policy":"https:\/\/doi.org\/10.1145\/crossmark-policy","source":"Crossref","is-referenced-by-count":14,"title":["ReLyMe"],"prefix":"10.1145","author":[{"given":"Chen","family":"Zhang","sequence":"first","affiliation":[{"name":"Zhejiang University, Hangzhou, China"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Luchin","family":"Chang","sequence":"additional","affiliation":[{"name":"Zhejiang University, Hangzhou, China"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Songruoyao","family":"Wu","sequence":"additional","affiliation":[{"name":"Zhejiang University, Hangzhou, China"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Xu","family":"Tan","sequence":"additional","affiliation":[{"name":"Microsoft Research Asia, Beijing, China"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Tao","family":"Qin","sequence":"additional","affiliation":[{"name":"Microsoft Research Asia, Beijing, China"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Tie-Yan","family":"Liu","sequence":"additional","affiliation":[{"name":"Microsoft Research Asia, Beijing, China"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Kejun","family":"Zhang","sequence":"additional","affiliation":[{"name":"Zhejiang University &amp; Alibaba-Zhejiang University Joint Institute of Frontier Technologies, Hangzhou, China"}],"role":[{"vocabulary":"crossref","role":"author"}]}],"member":"320","published-online":{"date-parts":[[2022,10,10]]},"reference":[{"key":"e_1_3_2_2_1_1","doi-asserted-by":"publisher","DOI":"10.1007\/978-3-030-32233-5_39"},{"key":"e_1_3_2_2_2_1","volume-title":"The materials of music: Sound and time. Music in theory and practice","author":"Benward B","year":"2003","unstructured":"B Benward and M Saker . 2003. Introduction . The materials of music: Sound and time. Music in theory and practice , pp. xii--25). NY, NY : McGraw-Hill ( 2003 ). B Benward and M Saker. 2003. Introduction. The materials of music: Sound and time. Music in theory and practice, pp. xii--25). NY, NY: McGraw-Hill (2003)."},{"key":"e_1_3_2_2_3_1","unstructured":"David Ronald Bishop. 2012. Perfect shark music: How can the principle of prosody be used in contemporary song-writing? Ph. D. Dissertation. Wintec.  David Ronald Bishop. 2012. Perfect shark music: How can the principle of prosody be used in contemporary song-writing? Ph. D. Dissertation. Wintec."},{"key":"e_1_3_2_2_4_1","doi-asserted-by":"publisher","DOI":"10.18653\/v1\/P19-1285"},{"key":"e_1_3_2_2_5_1","volume-title":"Prosodic phrasing is central to language comprehension. Trends in cognitive sciences","author":"Frazier Lyn","year":"2006","unstructured":"Lyn Frazier , Katy Carlson , and Charles Clifton Jr. 2006. Prosodic phrasing is central to language comprehension. Trends in cognitive sciences , Vol. 10 , 6 ( 2006 ), 244--249. Lyn Frazier, Katy Carlson, and Charles Clifton Jr. 2006. Prosodic phrasing is central to language comprehension. Trends in cognitive sciences, Vol. 10, 6 (2006), 244--249."},{"key":"e_1_3_2_2_6_1","volume-title":"Proc. 7th Sound and Music Computing Conference (SMC). 299--302","author":"Fukayama Satoru","year":"2010","unstructured":"Satoru Fukayama , Kei Nakatsuma , Shinji Sako , Takuya Nishimoto , and Shigeki Sagayama . 2010 . Automatic song composition from the lyrics exploiting prosody of the Japanese language . In Proc. 7th Sound and Music Computing Conference (SMC). 299--302 . Satoru Fukayama, Kei Nakatsuma, Shinji Sako, Takuya Nishimoto, and Shigeki Sagayama. 2010. Automatic song composition from the lyrics exploiting prosody of the Japanese language. In Proc. 7th Sound and Music Computing Conference (SMC). 299--302."},{"key":"e_1_3_2_2_7_1","unstructured":"Satoru Fukayama Daisuke Saito and Shigeki Sagayama. 2012. Assistance for novice users on creating songs from japanese lyrics. In ICMC.  Satoru Fukayama Daisuke Saito and Shigeki Sagayama. 2012. Assistance for novice users on creating songs from japanese lyrics. In ICMC."},{"key":"e_1_3_2_2_8_1","doi-asserted-by":"publisher","DOI":"10.5555\/1870658.1870709"},{"key":"e_1_3_2_2_9_1","doi-asserted-by":"crossref","unstructured":"Fenfei Guo Chen Zhang Zhirui Zhang Qixin He Kejun Zhang Jun Xie and Jordan Boyd-Graber. 2022. Automatic Song Translation for Tonal Languages. (2022).  Fenfei Guo Chen Zhang Zhirui Zhang Qixin He Kejun Zhang Jun Xie and Jordan Boyd-Graber. 2022. Automatic Song Translation for Tonal Languages. (2022).","DOI":"10.18653\/v1\/2022.findings-acl.60"},{"key":"e_1_3_2_2_10_1","doi-asserted-by":"publisher","DOI":"10.18653\/v1\/P17-1141"},{"key":"e_1_3_2_2_11_1","volume-title":"Lexical tones in Mandarin sung words: A Phonetic and Psycholiguistic investigation. Master's thesis","author":"Fangzhou Hu.","unstructured":"Fangzhou Hu. 2017. Lexical tones in Mandarin sung words: A Phonetic and Psycholiguistic investigation. Master's thesis . Shanghai International Studies University . Fangzhou Hu. 2017. Lexical tones in Mandarin sung words: A Phonetic and Psycholiguistic investigation. Master's thesis. Shanghai International Studies University."},{"key":"e_1_3_2_2_12_1","doi-asserted-by":"publisher","DOI":"10.18653\/v1\/N19-1090"},{"key":"e_1_3_2_2_13_1","doi-asserted-by":"publisher","DOI":"10.1145\/3394171.3413671"},{"key":"e_1_3_2_2_14_1","volume-title":"Machine Learning for Music Discovery Workshop, ICML.","author":"Jhamtani Harsh","year":"2019","unstructured":"Harsh Jhamtani and Taylor Berg-Kirkpatrick . 2019 . Modeling self-repetition in music generation using generative adversarial networks . In Machine Learning for Music Discovery Workshop, ICML. Harsh Jhamtani and Taylor Berg-Kirkpatrick. 2019. Modeling self-repetition in music generation using generative adversarial networks. In Machine Learning for Music Discovery Workshop, ICML."},{"key":"e_1_3_2_2_15_1","volume-title":"TeleMelody: Lyric-to-Melody Generation with a Template-Based Two-Stage Method. arXiv preprint arXiv:2109.09617","author":"Ju Zeqian","year":"2021","unstructured":"Zeqian Ju , Peiling Lu , Xu Tan , Rui Wang , Chen Zhang , Songruoyao Wu , Kejun Zhang , Xiangyang Li , Tao Qin , and Tie-Yan Liu . 2021. TeleMelody: Lyric-to-Melody Generation with a Template-Based Two-Stage Method. arXiv preprint arXiv:2109.09617 ( 2021 ). Zeqian Ju, Peiling Lu, Xu Tan, Rui Wang, Chen Zhang, Songruoyao Wu, Kejun Zhang, Xiangyang Li, Tao Qin, and Tie-Yan Liu. 2021. TeleMelody: Lyric-to-Melody Generation with a Template-Based Two-Stage Method. arXiv preprint arXiv:2109.09617 (2021)."},{"key":"e_1_3_2_2_16_1","volume-title":"Controlling the Output Length of Neural Machine Translation. In 16th International Workshop on Spoken Language Translation.","author":"Lakew Surafel Melaku","year":"2019","unstructured":"Surafel Melaku Lakew , Mattia Antonino Di Gangi , and Marcello Federico . 2019 . Controlling the Output Length of Neural Machine Translation. In 16th International Workshop on Spoken Language Translation. Surafel Melaku Lakew, Mattia Antonino Di Gangi, and Marcello Federico. 2019. Controlling the Output Length of Neural Machine Translation. In 16th International Workshop on Spoken Language Translation."},{"key":"e_1_3_2_2_17_1","volume-title":"Proceedings of the 2019 Conference of the North American Chapter of the Association for Computational Linguistics (Demonstrations). 84--88","author":"Lee Hsin-Pei","year":"2019","unstructured":"Hsin-Pei Lee , Jhih-Sheng Fang , and Wei-Yun Ma . 2019 . iComposer: An automatic songwriting system for Chinese popular music . In Proceedings of the 2019 Conference of the North American Chapter of the Association for Computational Linguistics (Demonstrations). 84--88 . Hsin-Pei Lee, Jhih-Sheng Fang, and Wei-Yun Ma. 2019. iComposer: An automatic songwriting system for Chinese popular music. In Proceedings of the 2019 Conference of the North American Chapter of the Association for Computational Linguistics (Demonstrations). 84--88."},{"key":"e_1_3_2_2_18_1","unstructured":"Fred Lerdahl and Ray Jackendoff. 1983. A Generative Theory pf Tonal Music.  Fred Lerdahl and Ray Jackendoff. 1983. A Generative Theory pf Tonal Music."},{"key":"e_1_3_2_2_19_1","doi-asserted-by":"publisher","DOI":"10.18653\/v1\/2020.acl-main.68"},{"key":"e_1_3_2_2_20_1","doi-asserted-by":"publisher","DOI":"10.1109\/ICDE.2013.6544937"},{"key":"e_1_3_2_2_21_1","doi-asserted-by":"publisher","DOI":"10.1145\/3474085.3475502"},{"key":"e_1_3_2_2_22_1","unstructured":"Kristine Monteith Tony R Martinez and Dan Ventura. 2012. Automatic Generation of Melodic Accompaniments for Lyrics.. In ICCC. 87--94.  Kristine Monteith Tony R Martinez and Dan Ventura. 2012. Automatic Generation of Melodic Accompaniments for Lyrics.. In ICCC. 87--94."},{"key":"e_1_3_2_2_23_1","volume-title":"ISMIR 2009-Proceedings of the 11th International Society for Music Information Retrieval Conference. 471--476","author":"Nichols Eric","year":"2009","unstructured":"Eric Nichols , Dan Morris , Sumit Basu , and Christopher Raphael . 2009 . Relationships between lyrics and melody in popular music . In ISMIR 2009-Proceedings of the 11th International Society for Music Information Retrieval Conference. 471--476 . Eric Nichols, Dan Morris, Sumit Basu, and Christopher Raphael. 2009. Relationships between lyrics and melody in popular music. In ISMIR 2009-Proceedings of the 11th International Society for Music Information Retrieval Conference. 471--476."},{"key":"e_1_3_2_2_24_1","volume-title":"a method for subjective performance assessment of the quality of speech voice output devices","author":"Rec ITUT","year":"1994","unstructured":"ITUT Rec . 1994. P. 85. a method for subjective performance assessment of the quality of speech voice output devices . International Telecommunication Union , Geneva ( 1994 ). ITUT Rec. 1994. P. 85. a method for subjective performance assessment of the quality of speech voice output devices. International Telecommunication Union, Geneva (1994)."},{"key":"e_1_3_2_2_25_1","doi-asserted-by":"publisher","DOI":"10.18653\/v1\/W19-5210"},{"key":"e_1_3_2_2_26_1","volume-title":"The realization of tone in singing in Cantonese and Mandarin. Ph.,D. Dissertation","author":"Schellenberg Murray Henry","unstructured":"Murray Henry Schellenberg . 2013. The realization of tone in singing in Cantonese and Mandarin. Ph.,D. Dissertation . University of British Columbia . Murray Henry Schellenberg. 2013. The realization of tone in singing in Cantonese and Mandarin. Ph.,D. Dissertation. University of British Columbia."},{"key":"e_1_3_2_2_27_1","volume-title":"General Theory of Modern Chinese","author":"Shao Jingmin","unstructured":"Jingmin Shao . 2007. General Theory of Modern Chinese . Shanghai Education Publishing House . Jingmin Shao. 2007. General Theory of Modern Chinese. Shanghai Education Publishing House."},{"key":"e_1_3_2_2_28_1","doi-asserted-by":"publisher","DOI":"10.1609\/aaai.v35i15.17626"},{"key":"e_1_3_2_2_29_1","volume-title":"International Conference on Machine Learning. PMLR, 2256--2265","author":"Sohl-Dickstein Jascha","year":"2015","unstructured":"Jascha Sohl-Dickstein , Eric Weiss , Niru Maheswaranathan , and Surya Ganguli . 2015 . Deep unsupervised learning using nonequilibrium thermodynamics . In International Conference on Machine Learning. PMLR, 2256--2265 . Jascha Sohl-Dickstein, Eric Weiss, Niru Maheswaranathan, and Surya Ganguli. 2015. Deep unsupervised learning using nonequilibrium thermodynamics. In International Conference on Machine Learning. PMLR, 2256--2265."},{"key":"e_1_3_2_2_30_1","volume-title":"Attention is all you need. Advances in neural information processing systems","author":"Vaswani Ashish","year":"2017","unstructured":"Ashish Vaswani , Noam Shazeer , Niki Parmar , Jakob Uszkoreit , Llion Jones , Aidan N Gomez , \u0141ukasz Kaiser , and Illia Polosukhin . 2017. Attention is all you need. Advances in neural information processing systems , Vol. 30 ( 2017 ). Ashish Vaswani, Noam Shazeer, Niki Parmar, Jakob Uszkoreit, Llion Jones, Aidan N Gomez, \u0141ukasz Kaiser, and Illia Polosukhin. 2017. Attention is all you need. Advances in neural information processing systems, Vol. 30 (2017)."},{"key":"e_1_3_2_2_31_1","first-page":"902","article-title":"Tone pitch realization under different intonation conditions","volume":"40","author":"Wang Susanna","year":"2015","unstructured":"Susanna Wang , Doyon Ding , and Higashi Taku . 2015 . Tone pitch realization under different intonation conditions . Journal of Acoustics , Vol. 40 , 6 (2015), 902 -- 913 . Susanna Wang, Doyon Ding, and Higashi Taku. 2015. Tone pitch realization under different intonation conditions. Journal of Acoustics, Vol. 40, 6 (2015), 902--913.","journal-title":"Journal of Acoustics"},{"key":"e_1_3_2_2_32_1","doi-asserted-by":"crossref","unstructured":"Moira Yip. 2002. Tone. Cambridge University Press.  Moira Yip. 2002. Tone. Cambridge University Press.","DOI":"10.1017\/CBO9781139164559"},{"key":"e_1_3_2_2_33_1","volume-title":"A study of cadential word relations","author":"Huiyong Yu.","unstructured":"Huiyong Yu. 2008. A study of cadential word relations . Central Conservatory of Music Press . Huiyong Yu. 2008. A study of cadential word relations. Central Conservatory of Music Press."},{"key":"e_1_3_2_2_34_1","doi-asserted-by":"publisher","DOI":"10.1145\/3424116"},{"key":"e_1_3_2_2_35_1","volume-title":"PDAugment: Data Augmentation by Pitch and Duration Adjustments for Automatic Lyrics Transcription. arXiv preprint arXiv:2109.07940","author":"Zhang Chen","year":"2021","unstructured":"Chen Zhang , Jiaxing Yu , LuChin Chang , Xu Tan , Jiawei Chen , Tao Qin , and Kejun Zhang . 2021a. PDAugment: Data Augmentation by Pitch and Duration Adjustments for Automatic Lyrics Transcription. arXiv preprint arXiv:2109.07940 ( 2021 ). Chen Zhang, Jiaxing Yu, LuChin Chang, Xu Tan, Jiawei Chen, Tao Qin, and Kejun Zhang. 2021a. PDAugment: Data Augmentation by Pitch and Duration Adjustments for Automatic Lyrics Transcription. arXiv preprint arXiv:2109.07940 (2021)."},{"key":"e_1_3_2_2_36_1","volume-title":"Structure-Enhanced Pop Music Generation via Harmony-Aware Learning. arXiv preprint arXiv:2109.06441","author":"Zhang Xueyao","year":"2021","unstructured":"Xueyao Zhang , Jinchao Zhang , Yao Qiu , Li Wang , and Jie Zhou . 2021b. Structure-Enhanced Pop Music Generation via Harmony-Aware Learning. arXiv preprint arXiv:2109.06441 ( 2021 ). Xueyao Zhang, Jinchao Zhang, Yao Qiu, Li Wang, and Jie Zhou. 2021b. Structure-Enhanced Pop Music Generation via Harmony-Aware Learning. arXiv preprint arXiv:2109.06441 (2021)."},{"key":"e_1_3_2_2_37_1","doi-asserted-by":"publisher","DOI":"10.1145\/3447548.3467418"}],"event":{"name":"MM '22: The 30th ACM International Conference on Multimedia","location":"Lisboa Portugal","acronym":"MM '22","sponsor":["SIGMM ACM Special Interest Group on Multimedia"]},"container-title":["Proceedings of the 30th ACM International Conference on Multimedia"],"original-title":[],"link":[{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/3503161.3548357","content-type":"unspecified","content-version":"vor","intended-application":"text-mining"},{"URL":"https:\/\/dl.acm.org\/doi\/pdf\/10.1145\/3503161.3548357","content-type":"unspecified","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2025,6,17]],"date-time":"2025-06-17T19:00:44Z","timestamp":1750186844000},"score":1,"resource":{"primary":{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/3503161.3548357"}},"subtitle":["Improving Lyric-to-Melody Generation by Incorporating Lyric-Melody Relationships"],"short-title":[],"issued":{"date-parts":[[2022,10,10]]},"references-count":37,"alternative-id":["10.1145\/3503161.3548357","10.1145\/3503161"],"URL":"https:\/\/doi.org\/10.1145\/3503161.3548357","relation":{},"subject":[],"published":{"date-parts":[[2022,10,10]]},"assertion":[{"value":"2022-10-10","order":2,"name":"published","label":"Published","group":{"name":"publication_history","label":"Publication History"}}]}}