{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2025,11,28]],"date-time":"2025-11-28T12:32:42Z","timestamp":1764333162468,"version":"3.41.0"},"publisher-location":"New York, NY, USA","reference-count":22,"publisher":"ACM","license":[{"start":{"date-parts":[[2022,4,27]],"date-time":"2022-04-27T00:00:00Z","timestamp":1651017600000},"content-version":"vor","delay-in-days":0,"URL":"https:\/\/www.acm.org\/publications\/policies\/copyright_policy#Background"}],"funder":[{"DOI":"10.13039\/501100020963","name":"Moonshot Research and Development Program","doi-asserted-by":"publisher","award":["JPMJMS2012"],"award-info":[{"award-number":["JPMJMS2012"]}],"id":[{"id":"10.13039\/501100020963","id-type":"DOI","asserted-by":"publisher"}]}],"content-domain":{"domain":[],"crossmark-restriction":false},"short-container-title":[],"published-print":{"date-parts":[[2022,4,27]]},"DOI":"10.1145\/3491101.3519700","type":"proceedings-article","created":{"date-parts":[[2022,4,29]],"date-time":"2022-04-29T16:49:48Z","timestamp":1651250988000},"page":"1-6","source":"Crossref","is-referenced-by-count":4,"title":["DualVoice: A Speech Interaction Method Using Whisper-Voice as Commands"],"prefix":"10.1145","author":[{"given":"Jun","family":"Rekimoto","sequence":"first","affiliation":[{"name":"The University of Tokyo, Japan and Sony CSL Kyoto, Japan"}],"role":[{"role":"author","vocabulary":"crossref"}]}],"member":"320","published-online":{"date-parts":[[2022,4,28]]},"reference":[{"key":"e_1_3_2_2_1_1","volume-title":"wav2vec 2.0: A Framework for Self-Supervised Learning of Speech Representations. arXiv [cs.CL] (June","author":"Baevski Alexei","year":"2020","unstructured":"Alexei Baevski , Henry Zhou , Abdelrahman Mohamed , and Michael Auli . 2020. wav2vec 2.0: A Framework for Self-Supervised Learning of Speech Representations. arXiv [cs.CL] (June 2020 ). Alexei Baevski, Henry Zhou, Abdelrahman Mohamed, and Michael Auli. 2020. wav2vec 2.0: A Framework for Self-Supervised Learning of Speech Representations. arXiv [cs.CL] (June 2020)."},{"key":"e_1_3_2_2_2_1","volume-title":"wav2vec 2.0: A Framework for Self-Supervised Learning of Speech Representations. arXiv [cs.CL] (June","author":"Baevski Alexei","year":"2020","unstructured":"Alexei Baevski , Henry Zhou , Abdelrahman Mohamed , and Michael Auli . 2020. wav2vec 2.0: A Framework for Self-Supervised Learning of Speech Representations. arXiv [cs.CL] (June 2020 ). Alexei Baevski, Henry Zhou, Abdelrahman Mohamed, and Michael Auli. 2020. wav2vec 2.0: A Framework for Self-Supervised Learning of Speech Representations. arXiv [cs.CL] (June 2020)."},{"key":"e_1_3_2_2_3_1","volume-title":"End-to-end Whispered Speech Recognition with Frequency-weighted Approaches and Pseudo Whisper Pre-training. (May","author":"Chang Heng-Jui","year":"2020","unstructured":"Heng-Jui Chang , Alexander\u00a0 H Liu , Hung-Yi Lee , and Lin-Shan Lee . 2020. End-to-end Whispered Speech Recognition with Frequency-weighted Approaches and Pseudo Whisper Pre-training. (May 2020 ). arxiv:2005.01972\u00a0[cs.CL] Heng-Jui Chang, Alexander\u00a0H Liu, Hung-Yi Lee, and Lin-Shan Lee. 2020. End-to-end Whispered Speech Recognition with Frequency-weighted Approaches and Pseudo Whisper Pre-training. (May 2020). arxiv:2005.01972\u00a0[cs.CL]"},{"key":"e_1_3_2_2_4_1","volume-title":"Voice Conversion for Whispered Speech Synthesis. (Dec","author":"Cotescu Marius","year":"2019","unstructured":"Marius Cotescu , Thomas Drugman , Goeric Huybrechts , Jaime Lorenzo-Trueba , and Alexis Moinet . 2019. Voice Conversion for Whispered Speech Synthesis. (Dec . 2019 ). arxiv:1912.05289\u00a0[cs.SD] Marius Cotescu, Thomas Drugman, Goeric Huybrechts, Jaime Lorenzo-Trueba, and Alexis Moinet. 2019. Voice Conversion for Whispered Speech Synthesis. (Dec. 2019). arxiv:1912.05289\u00a0[cs.SD]"},{"doi-asserted-by":"publisher","key":"e_1_3_2_2_5_1","DOI":"10.1145\/3242587.3242603"},{"doi-asserted-by":"publisher","key":"e_1_3_2_2_6_1","DOI":"10.1109\/TASLP.2016.2580944"},{"key":"e_1_3_2_2_7_1","volume-title":"Proc. of Eurospeech \u201999(1999)","author":"Goto Masataka","year":"1999","unstructured":"Masataka Goto . 1999 . A Real-time Filled Pause Detection System for Spontaneous Speech Recognition . Proc. of Eurospeech \u201999(1999) . Masataka Goto. 1999. A Real-time Filled Pause Detection System for Spontaneous Speech Recognition. Proc. of Eurospeech \u201999(1999)."},{"key":"e_1_3_2_2_8_1","volume-title":"8th European Conference on Speech Communication and Technology, EUROSPEECH 2003 - INTERSPEECH 2003","author":"Goto Masataka","year":"2003","unstructured":"Masataka Goto , Yukihiro Omoto , Katunobu Itou , and Tetsunori Kobayashi . 2003 . Speech shift: direct speech-input-mode switching through intentional control of voice pitch . In 8th European Conference on Speech Communication and Technology, EUROSPEECH 2003 - INTERSPEECH 2003 , Geneva, Switzerland , September 1-4, 2003. Masataka Goto, Yukihiro Omoto, Katunobu Itou, and Tetsunori Kobayashi. 2003. Speech shift: direct speech-input-mode switching through intentional control of voice pitch. In 8th European Conference on Speech Communication and Technology, EUROSPEECH 2003 - INTERSPEECH 2003, Geneva, Switzerland, September 1-4, 2003."},{"doi-asserted-by":"publisher","key":"e_1_3_2_2_9_1","DOI":"10.1109\/TASLP.2017.2738559"},{"unstructured":"Amazon.com Inc.2018. How Alexa keeps getting smarter. https:\/\/www.aboutamazon.com\/devices\/how-alexa-keeps-getting-smarter  Amazon.com Inc.2018. How Alexa keeps getting smarter. https:\/\/www.aboutamazon.com\/devices\/how-alexa-keeps-getting-smarter","key":"e_1_3_2_2_10_1"},{"unstructured":"Google Inc.2020. Google Cloud Speech-to-Text. https:\/\/cloud.google.com\/speech-to-text.  Google Inc.2020. Google Cloud Speech-to-Text. https:\/\/cloud.google.com\/speech-to-text.","key":"e_1_3_2_2_11_1"},{"volume-title":"Computational differences between whispered and non-whispered speech. Ph.\u00a0D. Dissertation","author":"Lim Boon\u00a0Pang","unstructured":"Boon\u00a0Pang Lim . 2010. Computational differences between whispered and non-whispered speech. Ph.\u00a0D. Dissertation . University of Illinois Urbana-Champaign. Boon\u00a0Pang Lim. 2010. Computational differences between whispered and non-whispered speech. Ph.\u00a0D. Dissertation. University of Illinois Urbana-Champaign.","key":"e_1_3_2_2_12_1"},{"key":"e_1_3_2_2_13_1","volume-title":"Pay Attention to MLPs. (May","author":"Liu Hanxiao","year":"2021","unstructured":"Hanxiao Liu , Zihang Dai , David\u00a0 R So , and Quoc\u00a0 V Le. 2021. Pay Attention to MLPs. (May 2021 ). arxiv:2105.08050\u00a0[cs.LG] Hanxiao Liu, Zihang Dai, David\u00a0R So, and Quoc\u00a0V Le. 2021. Pay Attention to MLPs. (May 2021). arxiv:2105.08050\u00a0[cs.LG]"},{"key":"e_1_3_2_2_14_1","volume-title":"An introduction to tkinter. www.pythonware.com\/library\/tkinter\/introduction\/index.htm","author":"Lundh Fredrik","year":"1999","unstructured":"Fredrik Lundh . 1999. An introduction to tkinter. www.pythonware.com\/library\/tkinter\/introduction\/index.htm ( 1999 ). Fredrik Lundh. 1999. An introduction to tkinter. www.pythonware.com\/library\/tkinter\/introduction\/index.htm (1999)."},{"key":"e_1_3_2_2_15_1","volume-title":"DualBreath: Input Method Using Nasal and Mouth Breathing. In Augmented Humans Conference 2021","author":"Onishi Ryoya","year":"2021","unstructured":"Ryoya Onishi , Tao Morisaki , Shun Suzuki , Saya Mizutani , Takaaki Kamigaki , Masahiro Fujiwara , Yasutoshi Makino , and Hiroyuki Shinoda . 2021 . DualBreath: Input Method Using Nasal and Mouth Breathing. In Augmented Humans Conference 2021 ( Rovaniemi, Finland) (AHs\u201921). Association for Computing Machinery, New York, NY, USA, 283\u2013285. Ryoya Onishi, Tao Morisaki, Shun Suzuki, Saya Mizutani, Takaaki Kamigaki, Masahiro Fujiwara, Yasutoshi Makino, and Hiroyuki Shinoda. 2021. DualBreath: Input Method Using Nasal and Mouth Breathing. In Augmented Humans Conference 2021(Rovaniemi, Finland) (AHs\u201921). Association for Computing Machinery, New York, NY, USA, 283\u2013285."},{"doi-asserted-by":"publisher","key":"e_1_3_2_2_16_1","DOI":"10.1109\/ICASSP.2015.7178964"},{"unstructured":"Adam Paszke Sam Gross Soumith Chintala Gregory Chanan Edward Yang Zachary DeVito Zeming Lin Alban Desmaison Luca Antiga and Adam Lerer. 2017. Automatic differentiation in PyTorch. (2017).  Adam Paszke Sam Gross Soumith Chintala Gregory Chanan Edward Yang Zachary DeVito Zeming Lin Alban Desmaison Luca Antiga and Adam Lerer. 2017. Automatic differentiation in PyTorch. (2017).","key":"e_1_3_2_2_17_1"},{"key":"e_1_3_2_2_18_1","volume-title":"Emotion Recognition from Speech Using Wav2vec 2.0 Embeddings. (April","author":"Pepino Leonardo","year":"2021","unstructured":"Leonardo Pepino , Pablo Riera , and Luciana Ferrer . 2021. Emotion Recognition from Speech Using Wav2vec 2.0 Embeddings. (April 2021 ). arxiv:2104.03502\u00a0[cs.SD] Leonardo Pepino, Pablo Riera, and Luciana Ferrer. 2021. Emotion Recognition from Speech Using Wav2vec 2.0 Embeddings. (April 2021). arxiv:2104.03502\u00a0[cs.SD]"},{"doi-asserted-by":"publisher","key":"e_1_3_2_2_19_1","DOI":"10.1145\/3411764.3445687"},{"unstructured":"Phil Wang. 2021. pytorch-gMLP: Implementation of gMLP an all-MLP replacement for Transformers in Pytorch. https:\/\/github.com\/lucidrains\/g-mlp-pytorch.  Phil Wang. 2021. pytorch-gMLP: Implementation of gMLP an all-MLP replacement for Transformers in Pytorch. https:\/\/github.com\/lucidrains\/g-mlp-pytorch.","key":"e_1_3_2_2_20_1"},{"key":"e_1_3_2_2_21_1","volume-title":"HuggingFace\u2019s Transformers: State-of-the-art Natural Language Processing. (Oct","author":"Wolf Thomas","year":"2019","unstructured":"Thomas Wolf , Lysandre Debut , Victor Sanh , Julien Chaumond , Clement Delangue , Anthony Moi , Pierric Cistac , Tim Rault , R\u00e9mi Louf , Morgan Funtowicz , Joe Davison , Sam Shleifer , Patrick von Platen , Clara Ma , Yacine Jernite , Julien Plu , Canwen Xu , Teven Le\u00a0Scao , Sylvain Gugger , Mariama Drame , Quentin Lhoest , and Alexander\u00a0 M Rush . 2019. HuggingFace\u2019s Transformers: State-of-the-art Natural Language Processing. (Oct . 2019 ). arxiv:1910.03771\u00a0[cs.CL] Thomas Wolf, Lysandre Debut, Victor Sanh, Julien Chaumond, Clement Delangue, Anthony Moi, Pierric Cistac, Tim Rault, R\u00e9mi Louf, Morgan Funtowicz, Joe Davison, Sam Shleifer, Patrick von Platen, Clara Ma, Yacine Jernite, Julien Plu, Canwen Xu, Teven Le\u00a0Scao, Sylvain Gugger, Mariama Drame, Quentin Lhoest, and Alexander\u00a0M Rush. 2019. HuggingFace\u2019s Transformers: State-of-the-art Natural Language Processing. (Oct. 2019). arxiv:1910.03771\u00a0[cs.CL]"},{"unstructured":"Cheng Yi Jianzhong Wang Ning Cheng Shiyu Zhou and Bo Xu. 2020. Applying Wav2vec2.0 to Speech Recognition in Various Low-resource Languages. (Dec. 2020). arxiv:2012.12121\u00a0[cs.CL]  Cheng Yi Jianzhong Wang Ning Cheng Shiyu Zhou and Bo Xu. 2020. Applying Wav2vec2.0 to Speech Recognition in Various Low-resource Languages. (Dec. 2020). arxiv:2012.12121\u00a0[cs.CL]","key":"e_1_3_2_2_22_1"}],"event":{"sponsor":["SIGCHI ACM Special Interest Group on Computer-Human Interaction"],"acronym":"CHI '22","name":"CHI '22: CHI Conference on Human Factors in Computing Systems","location":"New Orleans LA USA"},"container-title":["CHI Conference on Human Factors in Computing Systems Extended Abstracts"],"original-title":[],"link":[{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/3491101.3519700","content-type":"unspecified","content-version":"vor","intended-application":"text-mining"},{"URL":"https:\/\/dl.acm.org\/doi\/pdf\/10.1145\/3491101.3519700","content-type":"application\/pdf","content-version":"vor","intended-application":"syndication"},{"URL":"https:\/\/dl.acm.org\/doi\/pdf\/10.1145\/3491101.3519700","content-type":"unspecified","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2025,6,17]],"date-time":"2025-06-17T19:30:59Z","timestamp":1750188659000},"score":1,"resource":{"primary":{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/3491101.3519700"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2022,4,27]]},"references-count":22,"alternative-id":["10.1145\/3491101.3519700","10.1145\/3491101"],"URL":"https:\/\/doi.org\/10.1145\/3491101.3519700","relation":{},"subject":[],"published":{"date-parts":[[2022,4,27]]}}}