{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2026,6,13]],"date-time":"2026-06-13T07:06:55Z","timestamp":1781334415579,"version":"3.54.1"},"reference-count":96,"publisher":"Association for Computing Machinery (ACM)","issue":"2","license":[{"start":{"date-parts":[[2024,5,13]],"date-time":"2024-05-13T00:00:00Z","timestamp":1715558400000},"content-version":"vor","delay-in-days":0,"URL":"https:\/\/www.acm.org\/publications\/policies\/copyright_policy#Background"}],"content-domain":{"domain":["dl.acm.org"],"crossmark-restriction":true},"short-container-title":["Proc. ACM Interact. Mob. Wearable Ubiquitous Technol."],"published-print":{"date-parts":[[2024,5,13]]},"abstract":"<jats:p>Spatial awareness, particularly awareness of distant environmental scenes known as vista-space, is crucial and contributes to the cognitive and aesthetic needs of People with Visual Impairments (PVI). In this work, through a formative study with PVIs, we establish the need for vista-space awareness amongst people with visual impairments, and the possible scenarios where this awareness would be helpful. We investigate the potential of existing sonification techniques as well as AI-based audio generative models to design sounds that can create awareness of vista-space scenes. Our first user study, consisting of a listening test with sighted participants as well as PVIs, suggests that current AI generative models for audio have the potential to produce sounds that are comparable to existing sonification techniques in communicating sonic objects and scenes in terms of their intuitiveness, and learnability. Furthermore, through a wizard-of-oz study with PVIs, we demonstrate the utility of AI-generated sounds as well as scene audio recordings as auditory icons to provide vista-scene awareness, in the contexts of navigation and leisure. This is the first step towards addressing the need for vista-space awareness and experience in PVIs.<\/jats:p>","DOI":"10.1145\/3659609","type":"journal-article","created":{"date-parts":[[2024,5,15]],"date-time":"2024-05-15T12:20:41Z","timestamp":1715775641000},"page":"1-32","update-policy":"https:\/\/doi.org\/10.1145\/crossmark-policy","source":"Crossref","is-referenced-by-count":3,"title":["SonicVista: Towards Creating Awareness of Distant Scenes through Sonification"],"prefix":"10.1145","volume":"8","author":[{"ORCID":"https:\/\/orcid.org\/0000-0003-1350-9095","authenticated-orcid":false,"given":"Chitralekha","family":"Gupta","sequence":"first","affiliation":[{"name":"School of Computing, National University of Singapore, Singapore"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"ORCID":"https:\/\/orcid.org\/0009-0004-6675-3459","authenticated-orcid":false,"given":"Shreyas","family":"Sridhar","sequence":"additional","affiliation":[{"name":"School of Computing, National University of Singapore, Singapore"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"ORCID":"https:\/\/orcid.org\/0000-0001-9595-498X","authenticated-orcid":false,"given":"Denys J.C.","family":"Matthies","sequence":"additional","affiliation":[{"name":"Technical University of Applied Sciences L\u00fcbeck, Fraunhofer IMTE, Germany"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"ORCID":"https:\/\/orcid.org\/0000-0002-0768-1019","authenticated-orcid":false,"given":"Christophe","family":"Jouffrais","sequence":"additional","affiliation":[{"name":"IPAL, CNRS, Singapore"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"ORCID":"https:\/\/orcid.org\/0000-0001-7441-5493","authenticated-orcid":false,"given":"Suranga","family":"Nanayakkara","sequence":"additional","affiliation":[{"name":"School of Computing, National University of Singapore, Singapore"}],"role":[{"vocabulary":"crossref","role":"author"}]}],"member":"320","published-online":{"date-parts":[[2024,5,15]]},"reference":[{"key":"e_1_2_1_1_1","first-page":"4794","article-title":"Spatial Knowledge via Auditory Information for Blind Individuals","volume":"22","author":"Afonso-Jaco Amandine","year":"2022","unstructured":"Amandine Afonso-Jaco and Brian FG Katz. 2022. Spatial Knowledge via Auditory Information for Blind Individuals: Spatial Cognition Studies and the Use of Audio-VR. Sensors 22, 13 (2022), 4794.","journal-title":"Spatial Cognition Studies and the Use of Audio-VR. Sensors"},{"key":"e_1_2_1_2_1","doi-asserted-by":"crossref","unstructured":"Nida Aziz Tony Stockman and Rebecca Stewart. 2019. An investigation into customisable automatically generated auditory route overviews for pre-navigation. (2019).","DOI":"10.21785\/icad2019.029"},{"key":"e_1_2_1_3_1","doi-asserted-by":"publisher","DOI":"10.1145\/3173574.3173650"},{"key":"e_1_2_1_4_1","doi-asserted-by":"publisher","DOI":"10.1145\/3290605.3300528"},{"key":"e_1_2_1_5_1","volume-title":"Recognition-by-components: a theory of human image understanding. Psychological review 94, 2","author":"Biederman Irving","year":"1987","unstructured":"Irving Biederman. 1987. Recognition-by-components: a theory of human image understanding. Psychological review 94, 2 (1987), 115."},{"key":"e_1_2_1_6_1","doi-asserted-by":"publisher","DOI":"10.1145\/3264904"},{"key":"e_1_2_1_7_1","doi-asserted-by":"publisher","DOI":"10.1145\/3432196"},{"key":"e_1_2_1_8_1","volume-title":"One size fits all? What counts as quality practice in (reflexive) thematic analysis? Qualitative research in psychology 18, 3","author":"Braun Virginia","year":"2021","unstructured":"Virginia Braun and Victoria Clarke. 2021. One size fits all? What counts as quality practice in (reflexive) thematic analysis? Qualitative research in psychology 18, 3 (2021), 328--352."},{"key":"e_1_2_1_9_1","doi-asserted-by":"publisher","DOI":"10.1145\/169059.169179"},{"key":"e_1_2_1_10_1","doi-asserted-by":"publisher","DOI":"10.1109\/ICEV50249.2020.9289667"},{"key":"e_1_2_1_11_1","doi-asserted-by":"publisher","DOI":"10.1145\/3569478"},{"key":"e_1_2_1_12_1","volume-title":"Foley sound synthesis at the dcase 2023 challenge. arXiv preprint arXiv:2304.12521","author":"Choi Keunwoo","year":"2023","unstructured":"Keunwoo Choi, Jaekwon Im, Laurie Heller, Brian McFee, Keisuke Imoto, Yuki Okamoto, Mathieu Lagrange, and Shinosuke Takamichi. 2023. Foley sound synthesis at the dcase 2023 challenge. arXiv preprint arXiv:2304.12521 (2023)."},{"key":"e_1_2_1_13_1","volume-title":"HYU Submission For The DCASE 2023 Task 7: Diffusion Probabilistic Model With Adversarial Training For Foley Sound Synthesis. Technical Report. Department of Electronic Engineering","author":"Choi Won-Gook","year":"2023","unstructured":"Won-Gook Choi and Joon-Hyuk Chang. 2023. HYU Submission For The DCASE 2023 Task 7: Diffusion Probabilistic Model With Adversarial Training For Foley Sound Synthesis. Technical Report. Department of Electronic Engineering, Hanyang University, Seoul, Republic of Korea, Department of Electronic Engineering, Hanyang University, Seoul, Republic of Korea."},{"key":"e_1_2_1_14_1","volume-title":"Thematic analysis. Qualitative psychology: A practical guide to research methods 3","author":"Clarke Victoria","year":"2015","unstructured":"Victoria Clarke, Virginia Braun, and Nikki Hayfield. 2015. Thematic analysis. Qualitative psychology: A practical guide to research methods 3 (2015), 222--248."},{"key":"e_1_2_1_15_1","doi-asserted-by":"publisher","DOI":"10.1145\/3382507.3418874"},{"key":"e_1_2_1_16_1","volume-title":"Proceedings of the 14th International Conference on Auditory Display","author":"Dingler Tilman","year":"2008","unstructured":"Tilman Dingler, Jeffrey Lindsay, Bruce N Walker, et al. 2008. Learnability of sound cues for environmental features: Auditory icons, earcons, spearcons, and speech. In Proceedings of the 14th International Conference on Auditory Display, Paris, France. 1--6."},{"key":"e_1_2_1_17_1","volume-title":"Adversarial audio synthesis. arXiv preprint arXiv:1802.04208","author":"Donahue Chris","year":"2018","unstructured":"Chris Donahue, Julian McAuley, and Miller Puckette. 2018. Adversarial audio synthesis. arXiv preprint arXiv:1802.04208 (2018)."},{"key":"e_1_2_1_18_1","doi-asserted-by":"publisher","DOI":"10.1371\/journal.pone.0082491"},{"key":"e_1_2_1_19_1","volume-title":"Accessible interactive maps for visually impaired users. Mobility of Visually Impaired People: Fundamentals and ICT Assistive Technologies","author":"Ducasse Julie","year":"2018","unstructured":"Julie Ducasse, Anke M Brock, and Christophe Jouffrais. 2018. Accessible interactive maps for visually impaired users. Mobility of Visually Impaired People: Fundamentals and ICT Assistive Technologies (2018), 537--584."},{"key":"e_1_2_1_20_1","doi-asserted-by":"publisher","DOI":"10.1109\/ICASSP49357.2023.10095889"},{"key":"e_1_2_1_21_1","volume-title":"GANSynth: Adversarial Neural Audio Synthesis. In International Conference on Learning Representations.","author":"Engel Jesse","year":"2019","unstructured":"Jesse Engel, Kumar Krishna Agrawal, Shuo Chen, Ishaan Gulrajani, Chris Donahue, and Adam Roberts. 2019. GANSynth: Adversarial Neural Audio Synthesis. In International Conference on Learning Representations."},{"key":"e_1_2_1_22_1","volume-title":"DDSP: Differentiable Digital Signal Processing. In International Conference on Learning Representations.","author":"Engel Jesse","year":"2020","unstructured":"Jesse Engel, Lamtharn (Hanoi) Hantrakul, Chenjie Gu, and Adam Roberts. 2020. DDSP: Differentiable Digital Signal Processing. In International Conference on Learning Representations."},{"key":"e_1_2_1_23_1","volume-title":"Rodwan Hashim Mohammed Fallatah, and Jawad Syed","author":"Mohammed Fallatah Rodwan Hashim","year":"2018","unstructured":"Rodwan Hashim Mohammed Fallatah, Jawad Syed, Rodwan Hashim Mohammed Fallatah, and Jawad Syed. 2018. A critical review of Maslow's hierarchy of needs. Employee motivation in Saudi Arabia: An investigation into the higher education sector (2018), 19--59."},{"key":"e_1_2_1_24_1","doi-asserted-by":"publisher","DOI":"10.1111\/j.1460-2466.1968.tb00070.x"},{"key":"e_1_2_1_25_1","doi-asserted-by":"publisher","DOI":"10.1145\/28189.1044809"},{"key":"e_1_2_1_26_1","doi-asserted-by":"publisher","DOI":"10.1145\/169059.169184"},{"key":"e_1_2_1_27_1","doi-asserted-by":"publisher","DOI":"10.2307\/1574154"},{"key":"e_1_2_1_28_1","doi-asserted-by":"publisher","DOI":"10.1145\/3191741"},{"key":"e_1_2_1_29_1","doi-asserted-by":"crossref","first-page":"1","DOI":"10.1145\/3264919","article-title":"FootNotes: Geo-referenced audio annotations for nonvisual exploration","volume":"2","author":"Gleason Cole","year":"2018","unstructured":"Cole Gleason, Alexander J Fiannaca, Melanie Kneisel, Edward Cutrell, and Meredith Ringel Morris. 2018. FootNotes: Geo-referenced audio annotations for nonvisual exploration. Proceedings of the ACM on Interactive, Mobile, Wearable and Ubiquitous Technologies 2, 3 (2018), 1--24.","journal-title":"Proceedings of the ACM on Interactive, Mobile, Wearable and Ubiquitous Technologies"},{"key":"e_1_2_1_30_1","volume-title":"How Many Test Users in a Usability Study? https:\/\/www.nngroup.com\/articles\/how-many-test-users\/. [Online","author":"Nielsen Norman Group","year":"2024","unstructured":"Nielsen Norman Group. 2012. How Many Test Users in a Usability Study? https:\/\/www.nngroup.com\/articles\/how-many-test-users\/. [Online; accessed 23-Jan-2024]."},{"key":"e_1_2_1_31_1","volume-title":"How to Conduct Usability Studies for Accessibility. https:\/\/media.nngroup.com\/media\/reports\/free\/How_to_Conduct_Usability_Studies_for_Accessibility.pdf. [Online","author":"Nielsen Norman Group","year":"2024","unstructured":"Nielsen Norman Group. 2021. How to Conduct Usability Studies for Accessibility. https:\/\/media.nngroup.com\/media\/reports\/free\/How_to_Conduct_Usability_Studies_for_Accessibility.pdf. [Online; accessed 23-Jan-2024]."},{"key":"e_1_2_1_32_1","volume-title":"Towards Controllable Audio Texture Morphing. In ICASSP 2023-2023 IEEE International Conference on Acoustics, Speech and Signal Processing (ICASSP). IEEE, 1--5.","author":"Gupta Chitralekha","year":"2023","unstructured":"Chitralekha Gupta, Purnima Kamath, Yize Wei, Zhuoyao Li, Suranga Nanayakkara, and Lonce Wyse. 2023. Towards Controllable Audio Texture Morphing. In ICASSP 2023-2023 IEEE International Conference on Acoustics, Speech and Signal Processing (ICASSP). IEEE, 1--5."},{"key":"e_1_2_1_33_1","doi-asserted-by":"publisher","DOI":"10.1145\/2470654.2481292"},{"key":"e_1_2_1_34_1","unstructured":"Thomas Hermann Andy Hunt John G Neuhoff et al. 2011. The sonification handbook. Vol. 1. Logos Verlag Berlin."},{"key":"e_1_2_1_35_1","volume-title":"Denoising diffusion probabilistic models. Advances in neural information processing systems 33","author":"Ho Jonathan","year":"2020","unstructured":"Jonathan Ho, Ajay Jain, and Pieter Abbeel. 2020. Denoising diffusion probabilistic models. Advances in neural information processing systems 33 (2020), 6840--6851."},{"key":"e_1_2_1_36_1","doi-asserted-by":"publisher","DOI":"10.3390\/s20113222"},{"key":"e_1_2_1_37_1","volume-title":"Assistive augmentation","author":"Huber Jochen","unstructured":"Jochen Huber, Roy Shilkrot, Pattie Maes, and Suranga Nanayakkara. 2018. Assistive augmentation. Springer."},{"key":"e_1_2_1_38_1","doi-asserted-by":"publisher","DOI":"10.5244\/C.35.336"},{"key":"e_1_2_1_39_1","doi-asserted-by":"publisher","DOI":"10.1145\/3579496"},{"key":"e_1_2_1_40_1","doi-asserted-by":"crossref","unstructured":"Tiger F Ji Brianna R Cochran and Yuhang Zhao. 2022. Demonstration of VRBubble: Enhancing Peripheral Avatar Awareness for People with Visual Impairments in Social Virtual Reality. In CHI Conference on Human Factors in Computing Systems Extended Abstracts. 1--6.","DOI":"10.1145\/3491101.3519657"},{"key":"e_1_2_1_41_1","doi-asserted-by":"publisher","DOI":"10.3390\/s21103558"},{"key":"e_1_2_1_42_1","doi-asserted-by":"publisher","DOI":"10.1145\/3432216"},{"key":"e_1_2_1_43_1","doi-asserted-by":"publisher","DOI":"10.1145\/3025453.3025899"},{"key":"e_1_2_1_44_1","volume-title":"Example-Based Framework for Perceptually Guided Audio Texture Generation. arXiv preprint arXiv:2308.11859","author":"Kamath Purnima","year":"2023","unstructured":"Purnima Kamath, Chitralekha Gupta, Lonce Wyse, and Suranga Nanayakkara. 2023. Example-Based Framework for Perceptually Guided Audio Texture Generation. arXiv preprint arXiv:2308.11859 (2023)."},{"key":"e_1_2_1_45_1","volume-title":"Chitralekha Gupta, Lonce Wyse, and Suranga Nanayakkara.","author":"Kamath Purnima","year":"2023","unstructured":"Purnima Kamath, Tasnim Nishat Islam, Chitralekha Gupta, Lonce Wyse, and Suranga Nanayakkara. 2023. DCASE Task-7: StyleGAN2- Based Foley Sound Synthesis. Technical Report. National University of Singapore, Singapore and Bangladesh University of Engineering and Technology, Bangladesh and Universitat Pompeu Fabra, Barcelona, Spain."},{"key":"e_1_2_1_46_1","doi-asserted-by":"publisher","DOI":"10.2307\/3680062"},{"key":"e_1_2_1_47_1","doi-asserted-by":"publisher","DOI":"10.3233\/TAD-2012-0344"},{"key":"e_1_2_1_48_1","doi-asserted-by":"publisher","DOI":"10.1007\/s10055-012-0213-6"},{"key":"e_1_2_1_49_1","doi-asserted-by":"publisher","DOI":"10.1145\/3411825"},{"key":"e_1_2_1_50_1","doi-asserted-by":"publisher","DOI":"10.1109\/RO-MAN53752.2022.9900717"},{"key":"e_1_2_1_51_1","volume-title":"Diffwave: A versatile diffusion model for audio synthesis. arXiv preprint arXiv:2009.09761","author":"Kong Zhifeng","year":"2020","unstructured":"Zhifeng Kong, Wei Ping, Jiaji Huang, Kexin Zhao, and Bryan Catanzaro. 2020. Diffwave: A versatile diffusion model for audio synthesis. arXiv preprint arXiv:2009.09761 (2020)."},{"key":"e_1_2_1_52_1","doi-asserted-by":"publisher","DOI":"10.1145\/3290607.3312944"},{"key":"e_1_2_1_53_1","volume-title":"International Conference on Learning Representations (ICLR).","author":"Kreuk Felix","year":"2023","unstructured":"Felix Kreuk, Gabriel Synnaeve, Adam Polyak, Uriel Singer, Alexandre D\u00e9fossez, Jade Copet, Devi Parikh, Yaniv Taigman, and Yossi Adi. 2023. Audiogen: Textually guided audio generation. In International Conference on Learning Representations (ICLR)."},{"key":"e_1_2_1_54_1","doi-asserted-by":"publisher","DOI":"10.1016\/j.eswa.2022.118720"},{"key":"e_1_2_1_55_1","doi-asserted-by":"publisher","DOI":"10.1145\/3491102.3517490"},{"key":"e_1_2_1_56_1","doi-asserted-by":"publisher","DOI":"10.1145\/3290605.3300436"},{"key":"e_1_2_1_57_1","volume-title":"Proceedings of ICML.","author":"Liu Haohe","year":"2023","unstructured":"Haohe Liu, Zehua Chen, Yi Yuan, Xinhao Mei, Xubo Liu, Danilo Mandic, Wenwu Wang, and Mark D Plumbley. 2023. Audioldm: Text-to-audio generation with latent diffusion models. In Proceedings of ICML."},{"key":"e_1_2_1_58_1","volume-title":"Motivation and personality","author":"Maslow Abraham H","unstructured":"Abraham H Maslow. 1975. Motivation and personality. Harper & Row."},{"key":"e_1_2_1_59_1","volume-title":"Maslow's hierarchy of needs. Simply psychology 1","author":"McLeod Saul","year":"2007","unstructured":"Saul McLeod. 2007. Maslow's hierarchy of needs. Simply psychology 1 (2007), 1--18."},{"key":"e_1_2_1_60_1","doi-asserted-by":"publisher","DOI":"10.1177\/016502548801100105"},{"key":"e_1_2_1_61_1","doi-asserted-by":"publisher","DOI":"10.1007\/3-540-57207-4_21"},{"key":"e_1_2_1_62_1","doi-asserted-by":"publisher","DOI":"10.1016\/B0-08-043076-7\/01459-5"},{"key":"e_1_2_1_63_1","volume-title":"Assistive technology for students with visual impairments and blindness. Assistive technologies for people with diverse abilities","author":"Mulloy Austin M","year":"2014","unstructured":"Austin M Mulloy, Cindy Gevarter, Megan Hopkins, Kevin S Sutherland, and Sathiyaprakash T Ramdoss. 2014. Assistive technology for students with visual impairments and blindness. Assistive technologies for people with diverse abilities (2014), 113--156."},{"key":"e_1_2_1_64_1","volume-title":"DrumGAN: Synthesis of Drum Sounds With Timbral Feature Conditioning Using Generative Adversarial Networks. In International Society for Music Information Retrieval Conference.","author":"Nistal Javier","year":"2020","unstructured":"Javier Nistal, Stefan Lattner, and Ga\u00ebl Richard. 2020. DrumGAN: Synthesis of Drum Sounds With Timbral Feature Conditioning Using Generative Adversarial Networks. In International Society for Music Information Retrieval Conference."},{"key":"e_1_2_1_65_1","volume-title":"Proceedings of the 13th international conference on auditory display. 274--279","author":"Palladino Dianne K","year":"2007","unstructured":"Dianne K Palladino and Bruce N Walker. 2007. Learning rates for auditory menus enhanced with spearcons versus earcons. In Proceedings of the 13th international conference on auditory display. 274--279."},{"key":"e_1_2_1_66_1","doi-asserted-by":"publisher","DOI":"10.3389\/feduc.2021.723816"},{"key":"e_1_2_1_67_1","doi-asserted-by":"publisher","DOI":"10.1145\/3332165.3347881"},{"key":"e_1_2_1_68_1","doi-asserted-by":"publisher","DOI":"10.1017\/S0373463300013084"},{"key":"e_1_2_1_69_1","doi-asserted-by":"publisher","DOI":"10.1016\/j.ijhcs.2013.12.008"},{"key":"e_1_2_1_70_1","doi-asserted-by":"publisher","DOI":"10.1145\/3308561.3353779"},{"key":"e_1_2_1_71_1","volume-title":"International conference on machine learning. PMLR, 8748--8763","author":"Radford Alec","year":"2021","unstructured":"Alec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh, Gabriel Goh, Sandhini Agarwal, Girish Sastry, Amanda Askell, Pamela Mishkin, Jack Clark, et al. 2021. Learning transferable visual models from natural language supervision. In International conference on machine learning. PMLR, 8748--8763."},{"key":"e_1_2_1_72_1","doi-asserted-by":"publisher","DOI":"10.1145\/3130958"},{"key":"e_1_2_1_73_1","doi-asserted-by":"publisher","DOI":"10.17743\/jaes.2014.0009"},{"key":"e_1_2_1_74_1","doi-asserted-by":"publisher","DOI":"10.1109\/MED.2014.6961320"},{"key":"e_1_2_1_75_1","doi-asserted-by":"publisher","DOI":"10.1145\/3173574.3173717"},{"key":"e_1_2_1_76_1","doi-asserted-by":"publisher","DOI":"10.1145\/1978942.1979268"},{"key":"e_1_2_1_77_1","doi-asserted-by":"publisher","DOI":"10.1109\/IJCNN.2019.8852040"},{"key":"e_1_2_1_78_1","first-page":"29","article-title":"Location Tracking using Google Geolocation API","volume":"1","author":"Sharma Monika","year":"2015","unstructured":"Monika Sharma and Sudha Morwal. 2015. Location Tracking using Google Geolocation API. Int. J. of Sci. Tech. & Eng 1, 11 (2015), 29--32.","journal-title":"Int. J. of Sci. Tech. & Eng"},{"key":"e_1_2_1_79_1","doi-asserted-by":"publisher","DOI":"10.1109\/ICASSP49357.2023.10096023"},{"key":"e_1_2_1_80_1","doi-asserted-by":"publisher","DOI":"10.1145\/2702123.2702421"},{"key":"e_1_2_1_81_1","first-page":"1415","article-title":"Maximum likelihood training of score-based diffusion models","volume":"34","author":"Song Yang","year":"2021","unstructured":"Yang Song, Conor Durkan, Iain Murray, and Stefano Ermon. 2021. Maximum likelihood training of score-based diffusion models. Advances in Neural Information Processing Systems 34 (2021), 1415--1428.","journal-title":"Advances in Neural Information Processing Systems"},{"key":"e_1_2_1_82_1","doi-asserted-by":"publisher","DOI":"10.1145\/3161416"},{"key":"e_1_2_1_83_1","doi-asserted-by":"publisher","DOI":"10.1145\/228347.228369"},{"key":"e_1_2_1_84_1","doi-asserted-by":"publisher","DOI":"10.1145\/3313831.3376267"},{"key":"e_1_2_1_85_1","doi-asserted-by":"publisher","DOI":"10.1145\/2750858.2807541"},{"key":"e_1_2_1_86_1","doi-asserted-by":"publisher","DOI":"10.1145\/1714458.1714459"},{"key":"e_1_2_1_87_1","volume-title":"Proceedings of the 12th International Conference on Auditory Display. 63--68","author":"Walker Bruce N","year":"2006","unstructured":"Bruce N Walker, Amanda Nance, and Jeffrey Lindsay. 2006. Spearcons (speech-based earcons) improve navigation performance in auditory menus.. In Proceedings of the 12th International Conference on Auditory Display. 63--68."},{"key":"e_1_2_1_88_1","volume-title":"Managing the care of patients who have visual impairment. Nursing times 100, 1","author":"Watkinson Sue","year":"2004","unstructured":"Sue Watkinson and Eileen Scott. 2004. Managing the care of patients who have visual impairment. Nursing times 100, 1 (2004), 40--42."},{"key":"e_1_2_1_89_1","unstructured":"Keith White et al. 1991. Training Program for Individuals Working with Older American Indians Who Are Blind or Visually Impaired. Training Manual. (1991)."},{"key":"e_1_2_1_90_1","doi-asserted-by":"publisher","DOI":"10.1007\/978-3-031-03789-4_20"},{"key":"e_1_2_1_91_1","first-page":"1","article-title":"Virtual Paving: Rendering a smooth path for people with visual impairment through vibrotactile and audio feedback","volume":"4","author":"Xu Shuchang","year":"2020","unstructured":"Shuchang Xu, Ciyuan Yang, Wenhao Ge, Chun Yu, and Yuanchun Shi. 2020. Virtual Paving: Rendering a smooth path for people with visual impairment through vibrotactile and audio feedback. Proceedings of the ACM on Interactive, Mobile, Wearable and Ubiquitous Technologies 4, 3 (2020), 1--25.","journal-title":"Proceedings of the ACM on Interactive, Mobile, Wearable and Ubiquitous Technologies"},{"key":"e_1_2_1_92_1","doi-asserted-by":"publisher","DOI":"10.1145\/3495003"},{"key":"e_1_2_1_93_1","volume-title":"Diffsound: Discrete diffusion model for text-to-sound generation","author":"Yang Dongchao","year":"2023","unstructured":"Dongchao Yang, Jianwei Yu, Helin Wang, Wen Wang, Chao Weng, Yuexian Zou, and Dong Yu. 2023. Diffsound: Discrete diffusion model for text-to-sound generation. IEEE\/ACM Transactions on Audio, Speech, and Language Processing (2023)."},{"key":"e_1_2_1_94_1","doi-asserted-by":"publisher","DOI":"10.1145\/2559206.2581357"},{"key":"e_1_2_1_95_1","doi-asserted-by":"publisher","DOI":"10.1145\/3173574.3173690"},{"key":"e_1_2_1_96_1","doi-asserted-by":"publisher","DOI":"10.1145\/3173574.3173789"}],"container-title":["Proceedings of the ACM on Interactive, Mobile, Wearable and Ubiquitous Technologies"],"original-title":[],"language":"en","link":[{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/3659609","content-type":"unspecified","content-version":"vor","intended-application":"text-mining"},{"URL":"https:\/\/dl.acm.org\/doi\/pdf\/10.1145\/3659609","content-type":"unspecified","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2025,8,22]],"date-time":"2025-08-22T17:00:38Z","timestamp":1755882038000},"score":1,"resource":{"primary":{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/3659609"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2024,5,13]]},"references-count":96,"journal-issue":{"issue":"2","published-print":{"date-parts":[[2024,5,13]]}},"alternative-id":["10.1145\/3659609"],"URL":"https:\/\/doi.org\/10.1145\/3659609","relation":{},"ISSN":["2474-9567"],"issn-type":[{"value":"2474-9567","type":"electronic"}],"subject":[],"published":{"date-parts":[[2024,5,13]]},"assertion":[{"value":"2024-05-15","order":3,"name":"published","label":"Published","group":{"name":"publication_history","label":"Publication History"}}]}}