{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2026,2,21]],"date-time":"2026-02-21T13:19:48Z","timestamp":1771679988703,"version":"3.50.1"},"publisher-location":"New York, NY, USA","reference-count":20,"publisher":"ACM","license":[{"start":{"date-parts":[[2013,12,9]],"date-time":"2013-12-09T00:00:00Z","timestamp":1386547200000},"content-version":"vor","delay-in-days":0,"URL":"https:\/\/www.acm.org\/publications\/policies\/copyright_policy#Background"}],"content-domain":{"domain":["dl.acm.org"],"crossmark-restriction":true},"short-container-title":[],"published-print":{"date-parts":[[2013,12,9]]},"DOI":"10.1145\/2522848.2522856","type":"proceedings-article","created":{"date-parts":[[2013,11,27]],"date-time":"2013-11-27T14:14:02Z","timestamp":1385561642000},"page":"79-86","update-policy":"https:\/\/doi.org\/10.1145\/crossmark-policy","source":"Crossref","is-referenced-by-count":35,"title":["Predicting next speaker and timing from gaze transition patterns in multi-party meetings"],"prefix":"10.1145","author":[{"given":"Ryo","family":"Ishii","sequence":"first","affiliation":[{"name":"NTT Corporation, Kanagawa, Japan"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Kazuhiro","family":"Otsuka","sequence":"additional","affiliation":[{"name":"NTT Corporation, Kanagawa, Japan"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Shiro","family":"Kumano","sequence":"additional","affiliation":[{"name":"NTT Corporation, Kanagawa, Japan"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Masafumi","family":"Matsuda","sequence":"additional","affiliation":[{"name":"NTT Corporation, Kyoto, Japan"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Junji","family":"Yamato","sequence":"additional","affiliation":[{"name":"NTT Corporation, Kanagawa, Japan"}],"role":[{"role":"author","vocabulary":"crossref"}]}],"member":"320","published-online":{"date-parts":[[2013,12,9]]},"reference":[{"key":"e_1_3_2_1_1_1","doi-asserted-by":"publisher","DOI":"10.1145\/1647314.1647320"},{"key":"e_1_3_2_1_2_1","doi-asserted-by":"publisher","DOI":"10.1145\/1647314.1647332"},{"key":"e_1_3_2_1_3_1","doi-asserted-by":"crossref","first-page":"2306","DOI":"10.21437\/Interspeech.2010-632","volume-title":"INTERSPEECH","author":"Dielmann A.","year":"2010","unstructured":"A. Dielmann , G. Garau , and H. Bourlard . Floor holder detection and end of speaker turn prediction in meetings . In INTERSPEECH , pages 2306 -- 2309 , 2010 . A. Dielmann, G. Garau, and H. Bourlard. Floor holder detection and end of speaker turn prediction in meetings. In INTERSPEECH, pages 2306--2309, 2010."},{"key":"e_1_3_2_1_4_1","doi-asserted-by":"publisher","DOI":"10.1037\/h0033031"},{"key":"e_1_3_2_1_5_1","first-page":"2061","volume-title":"International Conference on Spoken Language Processing","volume":"3","author":"Ferrer L.","year":"2002","unstructured":"L. Ferrer , E. Shriberg , and A. Stolcke . Is the speaker done yet? faster and more accurate end-of-utterance detection using prosody in human-computer dialog . In International Conference on Spoken Language Processing , volume 3 , pages 2061 -- 2064 , 2002 . L. Ferrer, E. Shriberg, and A. Stolcke. Is the speaker done yet? faster and more accurate end-of-utterance detection using prosody in human-computer dialog. In International Conference on Spoken Language Processing, volume 3, pages 2061--2064, 2002."},{"key":"e_1_3_2_1_6_1","doi-asserted-by":"publisher","DOI":"10.1109\/MFI.2006.265658"},{"key":"e_1_3_2_1_7_1","first-page":"205","volume-title":"Biometrics","volume":"29","author":"Haberman S. J.","year":"1973","unstructured":"S. J. Haberman . The analysis of residuals in cross-classified tables . In Biometrics , volume 29 , pages 205 -- 220 , 1973 . S. J. Haberman. The analysis of residuals in cross-classified tables. In Biometrics, volume 29, pages 205--220, 1973."},{"key":"e_1_3_2_1_8_1","volume-title":"International Conference on Autonomous Agents and Multi-Agent Systems (AAMAS)","author":"Huang L.","year":"2011","unstructured":"L. Huang , L.-P. Morency , and J. Gratch . A multimodal end-of-turn prediction model: Learning from para social consensus sampling . In International Conference on Autonomous Agents and Multi-Agent Systems (AAMAS) , 2011 . L. Huang, L.-P. Morency, and J. Gratch. A multimodal end-of-turn prediction model: Learning from para social consensus sampling. In International Conference on Autonomous Agents and Multi-Agent Systems (AAMAS), 2011."},{"key":"e_1_3_2_1_9_1","first-page":"2018","volume-title":"INTERSPEECH","author":"Jokinen K.","year":"2011","unstructured":"K. Jokinen , K. Harada , M. Nishida , and S. Yamamoto . Turn-alignment using eye-gaze and speech in conversational interaction . In INTERSPEECH , pages 2018 -- 2021 , 2011 . K. Jokinen, K. Harada, M. Nishida, and S. Yamamoto. Turn-alignment using eye-gaze and speech in conversational interaction. In INTERSPEECH, pages 2018--2021, 2011."},{"key":"e_1_3_2_1_10_1","volume-title":"Conference of the European Chapter of the ACL","author":"Jovanovic N.","year":"2006","unstructured":"N. Jovanovic , R. op den Akker, and A. Nijholt. Addressee identification in face-to-face meetings . In Conference of the European Chapter of the ACL , 2006 . N. Jovanovic, R. op den Akker, and A. Nijholt. Addressee identification in face-to-face meetings. In Conference of the European Chapter of the ACL, 2006."},{"key":"e_1_3_2_1_11_1","volume-title":"INTERSPEECH","author":"Kawahara T.","year":"2012","unstructured":"T. Kawahara , T. Iwatate , and K. Takanashii . Prediction of turn-taking by combining prosodic and eye-gaze information in poster conversations . In INTERSPEECH , 2012 . T. Kawahara, T. Iwatate, and K. Takanashii. Prediction of turn-taking by combining prosodic and eye-gaze information in poster conversations. In INTERSPEECH, 2012."},{"key":"e_1_3_2_1_12_1","first-page":"22","article-title":"Some functions of gaze direction in social interaction","volume":"26","author":"Kendon A.","year":"1967","unstructured":"A. Kendon . Some functions of gaze direction in social interaction . ActaPsychologica , 26 : 22 -- 63 , 1967 . A. Kendon. Some functions of gaze direction in social interaction. ActaPsychologica, 26:22--63, 1967.","journal-title":"ActaPsychologica"},{"key":"e_1_3_2_1_13_1","first-page":"1367","volume-title":"INTERSPEECH","author":"Kipp M.","year":"2001","unstructured":"M. Kipp . Anvil - a generic annotation tool for multimodal dialogue . In INTERSPEECH , pages 1367 -- 1370 , 2001 . M. Kipp. Anvil - a generic annotation tool for multimodal dialogue. In INTERSPEECH, pages 1367--1370, 2001."},{"key":"e_1_3_2_1_14_1","first-page":"295","volume-title":"Language and Speech","volume":"41","author":"Koiso H.","year":"1998","unstructured":"H. Koiso , Y. Horiuchi , S. Tutiya , A. Ichikawa , and Y. Den . An analysis of turn-taking and backchannels based on prosodic and syntactic features in japanese map task dialogs . In Language and Speech , volume 41 , pages 295 -- 321 , 1998 . H. Koiso, Y. Horiuchi, S. Tutiya, A. Ichikawa, and Y. Den. An analysis of turn-taking and backchannels based on prosodic and syntactic features in japanese map task dialogs. In Language and Speech, volume 41, pages 295--321, 1998."},{"key":"e_1_3_2_1_15_1","doi-asserted-by":"publisher","DOI":"10.1109\/ICASSP.2011.5947629"},{"key":"e_1_3_2_1_16_1","volume-title":"SIGHAN","author":"Levow G.-A.","year":"2005","unstructured":"G.-A. Levow . Turn-taking in mandarin dialogue: Interactions of tones and intonation . In SIGHAN , 2005 . G.-A. Levow. Turn-taking in mandarin dialogue: Interactions of tones and intonation. In SIGHAN, 2005."},{"key":"e_1_3_2_1_17_1","doi-asserted-by":"publisher","DOI":"10.1007\/978-3-540-85483-8_18"},{"key":"e_1_3_2_1_18_1","doi-asserted-by":"publisher","DOI":"10.1109\/MSP.2011.941100"},{"key":"e_1_3_2_1_19_1","doi-asserted-by":"publisher","DOI":"10.1353\/lan.1974.0010"},{"key":"e_1_3_2_1_20_1","first-page":"17","volume-title":"INTERSPEECH","author":"Schlangen D.","year":"2006","unstructured":"D. Schlangen . From reaction to prediction experiments with computational models of turn-taking . In INTERSPEECH , pages 17 -- 21 , 2006 . D. Schlangen. From reaction to prediction experiments with computational models of turn-taking. In INTERSPEECH, pages 17--21, 2006."}],"event":{"name":"ICMI '13: 2013 International Conference on Multimodal Interaction","location":"Sydney Australia","acronym":"ICMI '13","sponsor":["SIGCHI ACM Special Interest Group on Computer-Human Interaction"]},"container-title":["Proceedings of the 15th ACM on International conference on multimodal interaction"],"original-title":[],"link":[{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/2522848.2522856","content-type":"unspecified","content-version":"vor","intended-application":"text-mining"},{"URL":"https:\/\/dl.acm.org\/doi\/pdf\/10.1145\/2522848.2522856","content-type":"unspecified","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2025,6,18]],"date-time":"2025-06-18T08:10:01Z","timestamp":1750234201000},"score":1,"resource":{"primary":{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/2522848.2522856"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2013,12,9]]},"references-count":20,"alternative-id":["10.1145\/2522848.2522856","10.1145\/2522848"],"URL":"https:\/\/doi.org\/10.1145\/2522848.2522856","relation":{},"subject":[],"published":{"date-parts":[[2013,12,9]]},"assertion":[{"value":"2013-12-09","order":2,"name":"published","label":"Published","group":{"name":"publication_history","label":"Publication History"}}]}}