{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2026,4,13]],"date-time":"2026-04-13T04:43:07Z","timestamp":1776055387266,"version":"3.50.1"},"publisher-location":"New York, NY, USA","reference-count":41,"publisher":"ACM","funder":[{"name":"the Research Grants Council of the Hong Kong Special Administrative Region","award":["PF23-94000"],"award-info":[{"award-number":["PF23-94000"]}]}],"content-domain":{"domain":["dl.acm.org"],"crossmark-restriction":true},"short-container-title":[],"published-print":{"date-parts":[[2026,4,13]]},"DOI":"10.1145\/3772363.3798425","type":"proceedings-article","created":{"date-parts":[[2026,4,13]],"date-time":"2026-04-13T01:55:28Z","timestamp":1776045328000},"page":"1-7","update-policy":"https:\/\/doi.org\/10.1145\/crossmark-policy","source":"Crossref","is-referenced-by-count":0,"title":["Sonic Stage: Automatically Generating an Interactive Spatial Soundscape to Facilitate Dialogue Video Comprehension for Blind and Low Vision Viewers"],"prefix":"10.1145","author":[{"ORCID":"https:\/\/orcid.org\/0000-0002-7642-9044","authenticated-orcid":false,"given":"Shuchang","family":"Xu","sequence":"first","affiliation":[{"name":"Computer Science and Engineering, The Hong Kong University of Science and Technology, Hong Kong, China"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"ORCID":"https:\/\/orcid.org\/0000-0002-7239-3769","authenticated-orcid":false,"given":"Xiaofu","family":"Jin","sequence":"additional","affiliation":[{"name":"Aalto University, Helsinki, Finland"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"ORCID":"https:\/\/orcid.org\/0000-0002-9084-9395","authenticated-orcid":false,"given":"Gaurav","family":"Jain","sequence":"additional","affiliation":[{"name":"Columbia University, New York, New York, USA"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"ORCID":"https:\/\/orcid.org\/0009-0007-9226-0713","authenticated-orcid":false,"given":"Wenshuo","family":"Zhang","sequence":"additional","affiliation":[{"name":"Computer Science and Engineering, The Hong Kong University of Science and Technology, Hong Kong, China"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"ORCID":"https:\/\/orcid.org\/0000-0002-3344-9694","authenticated-orcid":false,"given":"Huamin","family":"Qu","sequence":"additional","affiliation":[{"name":"Computer Science and Engineering, Hong Kong University of Science and Technology, Hong Kong, China"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"ORCID":"https:\/\/orcid.org\/0000-0003-2540-0839","authenticated-orcid":false,"given":"Brian A.","family":"Smith","sequence":"additional","affiliation":[{"name":"Columbia University, New York, New York, USA"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"ORCID":"https:\/\/orcid.org\/0000-0001-7515-3755","authenticated-orcid":false,"given":"Yukang","family":"Yan","sequence":"additional","affiliation":[{"name":"Department of Computer Science, University of Rochester, Rochester, New York, USA"}],"role":[{"role":"author","vocabulary":"crossref"}]}],"member":"320","published-online":{"date-parts":[[2026,4,13]]},"reference":[{"key":"e_1_3_3_3_2_2","doi-asserted-by":"publisher","DOI":"10.4324\/9780203892336"},{"key":"e_1_3_3_3_3_2","doi-asserted-by":"crossref","unstructured":"Virginia Braun and Victoria Clarke. 2019. Reflecting on reflexive thematic analysis. Qualitative research in sport exercise and health 11 4 (2019) 589\u2013597.","DOI":"10.1080\/2159676X.2019.1628806"},{"key":"e_1_3_3_3_4_2","doi-asserted-by":"crossref","unstructured":"Rick Busselle and Helena Bilandzic. 2009. Measuring narrative engagement. Media psychology 12 4 (2009) 321\u2013347.","DOI":"10.1080\/15213260903287259"},{"key":"e_1_3_3_3_5_2","doi-asserted-by":"publisher","DOI":"10.1145\/3126594.3126644"},{"key":"e_1_3_3_3_6_2","doi-asserted-by":"publisher","DOI":"10.1145\/3526113.3545613"},{"key":"e_1_3_3_3_7_2","doi-asserted-by":"crossref","unstructured":"Maryam Cheema Sina Elahimanesh Samuel Martin Pooyan Fazli and Hasti Seifi. 2025. DescribePro: Collaborative Audio Description with Human-AI Interaction. arXiv preprint arXiv:https:\/\/arXiv.org\/abs\/2508.01092 (2025).","DOI":"10.1145\/3663547.3746320"},{"key":"e_1_3_3_3_8_2","doi-asserted-by":"publisher","DOI":"10.1145\/3715336.3735685"},{"key":"e_1_3_3_3_9_2","doi-asserted-by":"publisher","DOI":"10.1145\/3654777.3676424"},{"key":"e_1_3_3_3_10_2","unstructured":"Gheorghe Comanici Eric Bieber Mike Schaekermann Ice Pasupat Noveen Sachdeva Inderjit Dhillon Marcel Blistein Ori Ram Dan Zhang Evan Rosen et\u00a0al. 2025. Gemini 2.5: Pushing the frontier with advanced reasoning multimodality long context and next generation agentic capabilities. arXiv preprint arXiv:https:\/\/arXiv.org\/abs\/2507.06261 (2025)."},{"key":"e_1_3_3_3_11_2","first-page":"324","volume-title":"Proceedings of the IEEE\/CVF Conference on Computer Vision and Pattern Recognition","author":"Gao Ruohan","year":"2019","unstructured":"Ruohan Gao and Kristen Grauman. 2019. 2.5 d visual sound. In Proceedings of the IEEE\/CVF Conference on Computer Vision and Pattern Recognition. 324\u2013333."},{"key":"e_1_3_3_3_12_2","doi-asserted-by":"crossref","unstructured":"Tilo Hartmann Werner Wirth Holger Schramm Christoph Klimmt Peter Vorderer Andr\u00e9 Gysbers Saskia B\u00f6cking Niklas Ravaja Jari Laarni Timo Saari et\u00a0al. 2015. The spatial presence experience scale (SPES). Journal of Media Psychology (2015).","DOI":"10.1037\/t48962-000"},{"key":"e_1_3_3_3_13_2","doi-asserted-by":"publisher","DOI":"10.1145\/3586183.3606830"},{"key":"e_1_3_3_3_14_2","doi-asserted-by":"publisher","DOI":"10.1145\/3597638.3608381"},{"key":"e_1_3_3_3_15_2","doi-asserted-by":"publisher","DOI":"10.1145\/3715070.3749268"},{"key":"e_1_3_3_3_16_2","doi-asserted-by":"publisher","DOI":"10.1145\/3597638.3608425"},{"key":"e_1_3_3_3_17_2","doi-asserted-by":"publisher","DOI":"10.1145\/3706598.3714096"},{"key":"e_1_3_3_3_18_2","doi-asserted-by":"publisher","DOI":"10.1145\/3586183.3606823"},{"key":"e_1_3_3_3_19_2","doi-asserted-by":"publisher","DOI":"10.1145\/3411764.3445644"},{"key":"e_1_3_3_3_20_2","unstructured":"Xingyu Liu Biao Wang Wayne Zhang Ziqian Liao Ziwen Li Amy Pavel Xiang\u2019Anthony\u2019 Chen et\u00a0al. 2025. CoSight: Exploring Viewer Contributions to Online Video Accessibility Through Descriptive Commenting. arXiv preprint arXiv:https:\/\/arXiv.org\/abs\/2508.08582 (2025)."},{"key":"e_1_3_3_3_21_2","doi-asserted-by":"publisher","DOI":"10.1145\/3526113.3545703"},{"key":"e_1_3_3_3_22_2","doi-asserted-by":"crossref","unstructured":"Mariana Lopez Gavin Kearney and Kriszti\u00e1n Hofst\u00e4dter. 2022. Seeing films through sound: Sound design spatial audio and accessibility for visually impaired audiences. British Journal of Visual Impairment 40 2 (2022) 117\u2013144.","DOI":"10.1177\/0264619620935935"},{"key":"e_1_3_3_3_23_2","doi-asserted-by":"crossref","unstructured":"Mariana\u00a0Julieta Lopez Gavin Kearney and Krisztian Hofstadter. 2021. Enhancing audio description: Inclusive cinematic experiences through sound design. Journal of Audiovisual Translation 4 1 (2021) 157\u2013182.","DOI":"10.47476\/jat.v4i1.2021.154"},{"key":"e_1_3_3_3_24_2","unstructured":"Pedro Morgado Nuno Nvasconcelos Timothy Langlois and Oliver Wang. 2018. Self-supervised generation of spatial audio for 360 video. Advances in neural information processing systems 31 (2018)."},{"key":"e_1_3_3_3_25_2","doi-asserted-by":"publisher","DOI":"10.1145\/3663548.3675617"},{"key":"e_1_3_3_3_26_2","doi-asserted-by":"publisher","DOI":"10.1145\/3544548.3581023"},{"key":"e_1_3_3_3_27_2","doi-asserted-by":"publisher","DOI":"10.1145\/3613904.3642632"},{"key":"e_1_3_3_3_28_2","doi-asserted-by":"publisher","DOI":"10.1145\/3635636.3656189"},{"key":"e_1_3_3_3_29_2","doi-asserted-by":"publisher","DOI":"10.1145\/3379337.3415864"},{"key":"e_1_3_3_3_30_2","doi-asserted-by":"publisher","DOI":"10.1145\/3441852.3471234"},{"key":"e_1_3_3_3_31_2","doi-asserted-by":"publisher","DOI":"10.1145\/3597638.3608402"},{"key":"e_1_3_3_3_32_2","doi-asserted-by":"publisher","DOI":"10.1145\/3613904.3642839"},{"key":"e_1_3_3_3_33_2","doi-asserted-by":"crossref","unstructured":"Valentijn\u00a0T Visch Ed\u00a0S Tan and Dylan Molenaar. 2010. The emotional and cognitive effect of immersion in film viewing. Cognition and Emotion 24 8 (2010) 1439\u20131445.","DOI":"10.1080\/02699930903498186"},{"key":"e_1_3_3_3_34_2","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR52734.2025.00499"},{"key":"e_1_3_3_3_35_2","doi-asserted-by":"publisher","DOI":"10.1145\/3411764.3445347"},{"key":"e_1_3_3_3_36_2","doi-asserted-by":"publisher","DOI":"10.1145\/3654777.3676336"},{"key":"e_1_3_3_3_37_2","doi-asserted-by":"publisher","DOI":"10.1145\/3706598.3713496"},{"key":"e_1_3_3_3_38_2","doi-asserted-by":"publisher","unstructured":"Shuchang Xu Xiaofu Jin Wenshuo Zhang Huamin Qu and Yukang Yan. 2025. Branch Explorer: Leveraging Branching Narratives to Support Interactive 360\u00b0 Video Viewing for Blind and Low Vision Users. Article 48 (2025) 18\u00a0pages. 10.1145\/3746059.3747791","DOI":"10.1145\/3746059.3747791"},{"key":"e_1_3_3_3_39_2","doi-asserted-by":"crossref","unstructured":"Shuchang Xu Ciyuan Yang Wenhao Ge Chun Yu and Yuanchun Shi. 2020. Virtual Paving: Rendering a smooth path for people with visual impairment through vibrotactile and audio feedback. Proceedings of the ACM on Interactive Mobile Wearable and Ubiquitous Technologies 4 3 (2020) 1\u201325.","DOI":"10.1145\/3411814"},{"key":"e_1_3_3_3_40_2","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR46437.2021.01523"},{"key":"e_1_3_3_3_41_2","doi-asserted-by":"publisher","unstructured":"Ciyuan Yang Shuchang Xu Tianyu Yu Guanhong Liu Chun Yu and Yuanchun Shi. 2021. LightGuide: Directing Visually Impaired People along a Path Using Light Cues. Proc. ACM Interact. Mob. Wearable Ubiquitous Technol. 5 2 Article 84 (June 2021) 27\u00a0pages. 10.1145\/3463524","DOI":"10.1145\/3463524"},{"key":"e_1_3_3_3_42_2","doi-asserted-by":"publisher","DOI":"10.1007\/978-3-030-58610-2_4"}],"event":{"name":"CHI EA '26: Extended Abstracts of the 2026 CHI Conference on Human Factors in Computing Systems","location":"Barcelona , Spain","acronym":"CHI EA '26","sponsor":["SIGCHI ACM Special Interest Group on Computer-Human Interaction"]},"container-title":["Proceedings of the Extended Abstracts of the 2026 CHI Conference on Human Factors in Computing Systems"],"original-title":[],"link":[{"URL":"https:\/\/dl.acm.org\/doi\/pdf\/10.1145\/3772363.3798425","content-type":"unspecified","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2026,4,13]],"date-time":"2026-04-13T03:49:01Z","timestamp":1776052141000},"score":1,"resource":{"primary":{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/3772363.3798425"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2026,4,13]]},"references-count":41,"alternative-id":["10.1145\/3772363.3798425","10.1145\/3772363"],"URL":"https:\/\/doi.org\/10.1145\/3772363.3798425","relation":{},"subject":[],"published":{"date-parts":[[2026,4,13]]},"assertion":[{"value":"2026-04-13","order":3,"name":"published","label":"Published","group":{"name":"publication_history","label":"Publication History"}}]}}