{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2026,5,21]],"date-time":"2026-05-21T01:35:54Z","timestamp":1779327354298,"version":"3.51.4"},"reference-count":16,"publisher":"Wiley","issue":"3","license":[{"start":{"date-parts":[[2003,11,4]],"date-time":"2003-11-04T00:00:00Z","timestamp":1067904000000},"content-version":"vor","delay-in-days":64,"URL":"http:\/\/onlinelibrary.wiley.com\/termsAndConditions#vor"}],"content-domain":{"domain":["onlinelibrary.wiley.com"],"crossmark-restriction":true},"short-container-title":["Computer Graphics Forum"],"published-print":{"date-parts":[[2003,9]]},"abstract":"<jats:title>Abstract<\/jats:title><jats:p> <jats:italic>Visemes are visual counterpart of phonemes. Traditionally, the speech animation of 3D synthetic faces involvesextraction of visemes from input speech followed by the application of co\u2010articulation rules to generate realisticanimation. In this paper, we take a novel approach for speech animation \u2014 using visyllables, the visual counterpartof syllables. The approach results into a concatenative visyllable based speech animation system. The key contributionof this paper lies in two main areas. Firstly, we define a set of visyllable units for spoken English along withthe associated phonological rules for valid syllables. Based on these rules, we have implemented a syllabificationalgorithm that allows segmentation of a given phoneme stream into syllables and subsequently visyllables. Secondly,we have recorded the database of visyllables using a facial motion capture system. The recorded visyllableunits are post\u2010processed semi\u2010automatically to ensure continuity at the vowel boundaries of the visyllables. We defineeach visyllable in terms of the Facial Movement Parameters (FMP). The FMPs are obtained as a result of thestatistical analysis of the facial motion capture data. The FMPs allow a compact representation of the visyllables.Further, the FMPs also facilitate the formulation of rules for boundary matching and smoothing after concatenatingthe visyllables units. Ours is the first visyllable based speech animation system. The proposed technique iseasy to implement, effective for real\u2010time as well as non real\u2010time applications and results into realistic speechanimation.<\/jats:italic> <\/jats:p><jats:p>Categories and Subject Descriptors (according to ACM CCS): 1.3.7 [Computer Graphics]: Three\u2010Dimensional Graphics and Realism<\/jats:p>","DOI":"10.1111\/1467-8659.t01-2-00711","type":"journal-article","created":{"date-parts":[[2003,11,4]],"date-time":"2003-11-04T18:30:59Z","timestamp":1067970659000},"page":"631-639","update-policy":"https:\/\/doi.org\/10.1002\/crossmark_policy","source":"Crossref","is-referenced-by-count":32,"title":["Visyllable Based Speech Animation"],"prefix":"10.1111","volume":"22","author":[{"given":"Sumedha","family":"Kshirsagar","sequence":"first","affiliation":[],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Nadia","family":"Magnenat\u2010Thalmann","sequence":"additional","affiliation":[],"role":[{"role":"author","vocabulary":"crossref"}]}],"member":"311","published-online":{"date-parts":[[2003,11,4]]},"reference":[{"key":"e_1_2_9_2_2","doi-asserted-by":"publisher","DOI":"10.1145\/258734.258880"},{"key":"e_1_2_9_3_2","doi-asserted-by":"publisher","DOI":"10.1007\/978-4-431-66911-1_13"},{"key":"e_1_2_9_4_2","volume-title":"An Introduction to the Pronunciation of English","author":"Gimson A.C.","year":"1983"},{"key":"e_1_2_9_5_2","first-page":"421\u0168","volume-title":"Proceedings of the IEEE International Conference on Acoustics, Speech and Signal Processing","author":"Hamaker J.","year":"1998"},{"issue":"1","key":"e_1_2_9_6_2","first-page":"1","article-title":"\u201ciface: a 3d synthetic talking face\u201d","volume":"31","author":"Hong P.","year":"2001","journal-title":"International Journal of Image and Graphics"},{"issue":"2","key":"e_1_2_9_7_2","first-page":"2037","article-title":"\u201cTriphone based unit selection for concatenative visual speech synthesis\u201d","author":"Huang F.S.","year":"2002","journal-title":"Proceedings of IEEE International Conference on Acoustics, Speech, and Signal Processing"},{"key":"e_1_2_9_8_2","doi-asserted-by":"crossref","first-page":"1171","DOI":"10.21437\/Eurospeech.1997-16","volume-title":"Proceedings Eurospeech '97","author":"Jones R.J.","year":"1997"},{"key":"e_1_2_9_9_2","volume-title":"Masters Thesis, Indiana University","author":"Kahn D.","year":"1976"},{"key":"e_1_2_9_10_2","volume-title":"Facial Action Coding System: A Technique for the Measurement of Facial Movement","author":"Friesen E.","year":"1978"},{"key":"e_1_2_9_11_2","unstructured":"Specification of MPEG\u20104 standard Moving Picture Experts Group.http:\/\/www.cselt.it\/mpeg."},{"key":"e_1_2_9_12_2","doi-asserted-by":"publisher","DOI":"10.1109\/CA.2001.982373"},{"key":"e_1_2_9_13_2","volume-title":"Proceedings of the Third ESCA Workshop on Speech Synthesis","author":"Kiraz G.","year":"1998"},{"key":"e_1_2_9_14_2","first-page":"33","volume-title":"Deformable Avatars","author":"Kshirsagar S.","year":"2001"},{"key":"e_1_2_9_15_2","doi-asserted-by":"publisher","DOI":"10.1109\/CGI.2001.934656"},{"key":"e_1_2_9_16_2","doi-asserted-by":"publisher","DOI":"10.1207\/s15516709cog2001_1"},{"key":"e_1_2_9_17_2","unstructured":"Vicon motion capture system.http:\/\/www.vicon.com\/"}],"container-title":["Computer Graphics Forum"],"original-title":[],"language":"en","link":[{"URL":"https:\/\/api.wiley.com\/onlinelibrary\/tdm\/v1\/articles\/10.1111%2F1467-8659.t01-2-00711","content-type":"unspecified","content-version":"vor","intended-application":"text-mining"},{"URL":"https:\/\/onlinelibrary.wiley.com\/doi\/pdf\/10.1111\/1467-8659.t01-2-00711","content-type":"unspecified","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2023,10,8]],"date-time":"2023-10-08T18:11:26Z","timestamp":1696788686000},"score":1,"resource":{"primary":{"URL":"https:\/\/onlinelibrary.wiley.com\/doi\/10.1111\/1467-8659.t01-2-00711"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2003,9]]},"references-count":16,"journal-issue":{"issue":"3","published-print":{"date-parts":[[2003,9]]}},"alternative-id":["10.1111\/1467-8659.t01-2-00711"],"URL":"https:\/\/doi.org\/10.1111\/1467-8659.t01-2-00711","archive":["Portico"],"relation":{},"ISSN":["0167-7055","1467-8659"],"issn-type":[{"value":"0167-7055","type":"print"},{"value":"1467-8659","type":"electronic"}],"subject":[],"published":{"date-parts":[[2003,9]]},"assertion":[{"value":"2003-11-04","order":2,"name":"published","label":"Published","group":{"name":"publication_history","label":"Publication History"}}]}}