{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2025,6,18]],"date-time":"2025-06-18T04:24:38Z","timestamp":1750220678187,"version":"3.41.0"},"reference-count":55,"publisher":"Association for Computing Machinery (ACM)","issue":"6","license":[{"start":{"date-parts":[[2020,11,27]],"date-time":"2020-11-27T00:00:00Z","timestamp":1606435200000},"content-version":"vor","delay-in-days":0,"URL":"https:\/\/www.acm.org\/publications\/policies\/copyright_policy#Background"}],"content-domain":{"domain":["dl.acm.org"],"crossmark-restriction":true},"short-container-title":["ACM Trans. Graph."],"published-print":{"date-parts":[[2020,12,31]]},"abstract":"<jats:p>Audio recordings contain rich information about sound sources and their properties such as the location, loudness, and frequency of events. One prevalent component in sound recordings is the sound texture, which contains a massive number of events. In such a texture, there can be some distinct and repeated sounds that we term as a foreground sound. Birds chirping in the wind is one such decorative sound texture with the chirping as a foreground sound and the wind as a background texture. To render these decorative sound textures in real-time and with high quality, we create two-layer Markov Models to enable smooth transitions from sound grain to sound grain and propose a hierarchical scheme to generate Head-Related Transfer Function filters for localization cues of sounds represented as area\/volume sources. Moreover, during the synthesis stage, we provide control over the frequency and intensity of sounds for customization. Lastly, foreground sounds are often blended into background textures such as the sound of rain splats on car surfaces becoming submerged in the background rain. We develop an extraction component that outperforms existing learning-based methods to facilitate our synthesis with perceptible foreground sounds and well-defined textures.<\/jats:p>","DOI":"10.1145\/3414685.3417875","type":"journal-article","created":{"date-parts":[[2020,11,27]],"date-time":"2020-11-27T21:51:05Z","timestamp":1606513865000},"page":"1-12","update-policy":"https:\/\/doi.org\/10.1145\/crossmark-policy","source":"Crossref","is-referenced-by-count":2,"title":["Real-time rendering of decorative sound textures for soundscapes"],"prefix":"10.1145","volume":"39","author":[{"given":"Jinta","family":"Zheng","sequence":"first","affiliation":[{"name":"Oregon State University"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Shih-Hsuan","family":"Hung","sequence":"additional","affiliation":[{"name":"Oregon State University"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Kyle","family":"Hiebel","sequence":"additional","affiliation":[{"name":"Oregon State University"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Yue","family":"Zhang","sequence":"additional","affiliation":[{"name":"Oregon State University"}],"role":[{"role":"author","vocabulary":"crossref"}]}],"member":"320","published-online":{"date-parts":[[2020,11,27]]},"reference":[{"key":"e_1_2_2_1_1","volume-title":"IEEE, 2001 IEEE Workshop, 99--102","author":"Algazi V Ralph","year":"2001","unstructured":"V Ralph Algazi , Richard O Duda , Dennis M Thompson , and Carlos Avendano . 2001 . The CIPIC HRTF database. In Applications of Signal Processing to Audio and Acoustics . IEEE, 2001 IEEE Workshop, 99--102 . V Ralph Algazi, Richard O Duda, Dennis M Thompson, and Carlos Avendano. 2001. The CIPIC HRTF database. In Applications of Signal Processing to Audio and Acoustics. IEEE, 2001 IEEE Workshop, 99--102."},{"key":"e_1_2_2_2_1","unstructured":"Durand R Begault and Leonard J Trejo. 2000. 3-D sound for virtual reality and multimedia. (2000).  Durand R Begault and Leonard J Trejo. 2000. 3-D sound for virtual reality and multimedia. (2000)."},{"key":"e_1_2_2_3_1","doi-asserted-by":"publisher","DOI":"10.1109\/TSA.2005.851998"},{"key":"e_1_2_2_4_1","doi-asserted-by":"publisher","DOI":"10.1109\/TASSP.1979.1163209"},{"key":"e_1_2_2_5_1","volume-title":"Audio texture synthesis with scattering moments. arXiv preprint arXiv:1311.0407","author":"Bruna Joan","year":"2013","unstructured":"Joan Bruna and St\u00e9phane Mallat . 2013. Audio texture synthesis with scattering moments. arXiv preprint arXiv:1311.0407 ( 2013 ). Joan Bruna and St\u00e9phane Mallat. 2013. Audio texture synthesis with scattering moments. arXiv preprint arXiv:1311.0407 (2013)."},{"key":"e_1_2_2_6_1","volume-title":"International Conference on Machine Learning. 208--216","author":"Bryan Nicholas","year":"2013","unstructured":"Nicholas Bryan and Gautham Mysore . 2013 . An efficient posterior regularized latent variable model for interactive sound source separation . In International Conference on Machine Learning. 208--216 . Nicholas Bryan and Gautham Mysore. 2013. An efficient posterior regularized latent variable model for interactive sound source separation. In International Conference on Machine Learning. 208--216."},{"key":"e_1_2_2_7_1","first-page":"1","article-title":"Interactive sound propagation with bidirectional path tracing","volume":"35","author":"Cao Chunxiao","year":"2016","unstructured":"Chunxiao Cao , Zhong Ren , Carl Schissler , Dinesh Manocha , and Kun Zhou . 2016 . Interactive sound propagation with bidirectional path tracing . ACM Transactions on Graphics (TOG) 35 , 6 (2016), 1 -- 11 . Chunxiao Cao, Zhong Ren, Carl Schissler, Dinesh Manocha, and Kun Zhou. 2016. Interactive sound propagation with bidirectional path tracing. ACM Transactions on Graphics (TOG) 35, 6 (2016), 1--11.","journal-title":"ACM Transactions on Graphics (TOG)"},{"key":"e_1_2_2_8_1","doi-asserted-by":"publisher","DOI":"10.1145\/1964921.1964979"},{"key":"e_1_2_2_9_1","doi-asserted-by":"publisher","DOI":"10.1145\/3197517.3201371"},{"key":"e_1_2_2_10_1","doi-asserted-by":"publisher","DOI":"10.1016\/j.jspi.2009.05.038"},{"key":"e_1_2_2_11_1","volume-title":"CARLA: An open urban driving simulator. arXiv preprint arXiv:1711.03938","author":"Dosovitskiy Alexey","year":"2017","unstructured":"Alexey Dosovitskiy , German Ros , Felipe Codevilla , Antonio Lopez , and Vladlen Koltun . 2017 . CARLA: An open urban driving simulator. arXiv preprint arXiv:1711.03938 (2017). Alexey Dosovitskiy, German Ros, Felipe Codevilla, Antonio Lopez, and Vladlen Koltun. 2017. CARLA: An open urban driving simulator. arXiv preprint arXiv:1711.03938 (2017)."},{"key":"e_1_2_2_12_1","doi-asserted-by":"publisher","DOI":"10.1121\/1.423749"},{"key":"e_1_2_2_13_1","volume-title":"Finkel and Jon Louis Bentley","author":"Raphael","year":"1974","unstructured":"Raphael A. Finkel and Jon Louis Bentley . 1974 . Quad trees a data structure for retrieval on composite keys. Acta informatica 4, 1 (1974), 1--9. Raphael A. Finkel and Jon Louis Bentley. 1974. Quad trees a data structure for retrieval on composite keys. Acta informatica 4, 1 (1974), 1--9."},{"key":"e_1_2_2_14_1","doi-asserted-by":"publisher","DOI":"10.1145\/2502081.2502245"},{"key":"e_1_2_2_15_1","volume-title":"Audio Engineering Society Conference: 22nd International Conference: Virtual, Synthetic, and Entertainment Audio. Audio Engineering Society.","author":"Freeland Fabio P","year":"2002","unstructured":"Fabio P Freeland , Luiz WP Biscainho , and Paulo SR Diniz . 2002 . Efficient HRTF interpolation in 3D moving sound . In Audio Engineering Society Conference: 22nd International Conference: Virtual, Synthetic, and Entertainment Audio. Audio Engineering Society. Fabio P Freeland, Luiz WP Biscainho, and Paulo SR Diniz. 2002. Efficient HRTF interpolation in 3D moving sound. In Audio Engineering Society Conference: 22nd International Conference: Virtual, Synthetic, and Entertainment Audio. Audio Engineering Society."},{"key":"e_1_2_2_16_1","doi-asserted-by":"publisher","DOI":"10.1121\/1.4828983"},{"key":"e_1_2_2_17_1","first-page":"618","article-title":"Augmented reality audio for mobile and wearable appliances","volume":"52","author":"H\u00e4rm\u00e4 Aki","year":"2004","unstructured":"Aki H\u00e4rm\u00e4 , Julia Jakka , Miikka Tikander , Matti Karjalainen , Tapio Lokki , Jarmo Hiipakka , and Ga\u00ebtan Lorho . 2004 . Augmented reality audio for mobile and wearable appliances . Journal of the Audio Engineering Society 52 , 6 (2004), 618 -- 639 . Aki H\u00e4rm\u00e4, Julia Jakka, Miikka Tikander, Matti Karjalainen, Tapio Lokki, Jarmo Hiipakka, and Ga\u00ebtan Lorho. 2004. Augmented reality audio for mobile and wearable appliances. Journal of the Audio Engineering Society 52, 6 (2004), 618--639.","journal-title":"Journal of the Audio Engineering Society"},{"key":"e_1_2_2_18_1","doi-asserted-by":"publisher","DOI":"10.1186\/1687-4722-2014-9"},{"key":"e_1_2_2_19_1","volume-title":"Invariance to background noise as a signature of non-primary auditory cortex. Nature communications 10, 1","author":"Kell Alexander JE","year":"2019","unstructured":"Alexander JE Kell and Josh H McDermott . 2019. Invariance to background noise as a signature of non-primary auditory cortex. Nature communications 10, 1 ( 2019 ), 1--11. Alexander JE Kell and Josh H McDermott. 2019. Invariance to background noise as a signature of non-primary auditory cortex. Nature communications 10, 1 (2019), 1--11."},{"key":"e_1_2_2_20_1","doi-asserted-by":"publisher","DOI":"10.1145\/1186822.1073263"},{"key":"e_1_2_2_21_1","doi-asserted-by":"publisher","DOI":"10.1145\/882262.882264"},{"key":"e_1_2_2_22_1","volume-title":"Proc. of the 16th Int. Conference on Digital Audio Effects (DAFx-13)","author":"Liao Wei-Hsiang","year":"2013","unstructured":"Wei-Hsiang Liao , Axel Roebel , and Alvin Su . 2013 . On the modeling of sound textures based on the STFT representation . In Proc. of the 16th Int. Conference on Digital Audio Effects (DAFx-13) . 33. Wei-Hsiang Liao, Axel Roebel, and Alvin Su. 2013. On the modeling of sound textures based on the STFT representation. In Proc. of the 16th Int. Conference on Digital Audio Effects (DAFx-13). 33."},{"key":"e_1_2_2_23_1","doi-asserted-by":"publisher","DOI":"10.1145\/3306346.3323045"},{"key":"e_1_2_2_24_1","doi-asserted-by":"publisher","DOI":"10.1016\/j.neuron.2011.06.032"},{"key":"e_1_2_2_25_1","doi-asserted-by":"publisher","DOI":"10.1109\/TASLP.2018.2858559"},{"key":"e_1_2_2_26_1","doi-asserted-by":"publisher","DOI":"10.1109\/TASLP.2015.2389618"},{"key":"e_1_2_2_27_1","volume-title":"Voice recognition algorithms using mel frequency cepstral coefficient (MFCC) and dynamic time warping (DTW) techniques. arXiv preprint arXiv:1003.4083","author":"Muda Lindasalwa","year":"2010","unstructured":"Lindasalwa Muda , Mumtaj Begam , and Irraivan Elamvazuthi . 2010. Voice recognition algorithms using mel frequency cepstral coefficient (MFCC) and dynamic time warping (DTW) techniques. arXiv preprint arXiv:1003.4083 ( 2010 ). Lindasalwa Muda, Mumtaj Begam, and Irraivan Elamvazuthi. 2010. Voice recognition algorithms using mel frequency cepstral coefficient (MFCC) and dynamic time warping (DTW) techniques. arXiv preprint arXiv:1003.4083 (2010)."},{"key":"e_1_2_2_28_1","unstructured":"Sean O'Leary and Axel Roebel. 2014. A two level montage approach to sound texture synthesis with treatment of unique events.. In DAFx. 1--1.  Sean O'Leary and Axel Roebel. 2014. A two level montage approach to sound texture synthesis with treatment of unique events.. In DAFx. 1--1."},{"key":"e_1_2_2_29_1","doi-asserted-by":"publisher","DOI":"10.1109\/TASLP.2016.2536481"},{"key":"e_1_2_2_30_1","volume-title":"Psychoacoustic model compensation for robust speaker verification in environmental noise","author":"Panda Ashish","year":"2011","unstructured":"Ashish Panda and Thambipillai Srikanthan . 2011. Psychoacoustic model compensation for robust speaker verification in environmental noise . IEEE transactions on audio, speech, and language processing 20, 3 ( 2011 ), 945--953. Ashish Panda and Thambipillai Srikanthan. 2011. Psychoacoustic model compensation for robust speaker verification in environmental noise. IEEE transactions on audio, speech, and language processing 20, 3 (2011), 945--953."},{"key":"e_1_2_2_31_1","doi-asserted-by":"publisher","DOI":"10.1121\/1.399421"},{"key":"e_1_2_2_32_1","doi-asserted-by":"publisher","DOI":"10.1145\/344779.344987"},{"volume-title":"Theory and applications of digital speech processing","author":"Rabiner Lawrence R","key":"e_1_2_2_33_1","unstructured":"Lawrence R Rabiner and Ronald W Schafer . 2011. Theory and applications of digital speech processing . Vol. 64 . Pearson Upper Saddle River , NJ. Lawrence R Rabiner and Ronald W Schafer. 2011. Theory and applications of digital speech processing. Vol. 64. Pearson Upper Saddle River, NJ."},{"key":"e_1_2_2_34_1","doi-asserted-by":"publisher","DOI":"10.1121\/1.3278605"},{"key":"e_1_2_2_35_1","doi-asserted-by":"publisher","DOI":"10.1109\/TVCG.2009.27"},{"key":"e_1_2_2_36_1","doi-asserted-by":"publisher","DOI":"10.1145\/3197517.3201339"},{"key":"e_1_2_2_37_1","doi-asserted-by":"publisher","DOI":"10.2307\/3679937"},{"key":"e_1_2_2_38_1","doi-asserted-by":"publisher","DOI":"10.1109\/JSTSP.2015.2421876"},{"key":"e_1_2_2_39_1","unstructured":"Nicolas Saint-Arnaud and Kris Popat. 1995. Analysis and synthesis of sound textures. In in Readings in Computational Auditory Scene Analysis. Citeseer.  Nicolas Saint-Arnaud and Kris Popat. 1995. Analysis and synthesis of sound textures. In in Readings in Computational Auditory Scene Analysis. Citeseer."},{"key":"e_1_2_2_40_1","doi-asserted-by":"publisher","DOI":"10.1145\/2601097.2601216"},{"key":"e_1_2_2_41_1","doi-asserted-by":"publisher","DOI":"10.1109\/TVCG.2016.2518134"},{"key":"e_1_2_2_42_1","unstructured":"Diemo Schwarz. 2011. State of the art in sound texture synthesis. In Digital Audio Effects (DAFx). 221--232.  Diemo Schwarz. 2011. State of the art in sound texture synthesis. In Digital Audio Effects (DAFx). 221--232."},{"key":"e_1_2_2_43_1","volume-title":"International Symposium on Computer Music Multidisciplinary Research. Springer, 372--392","author":"Schwarz Diemo","year":"2013","unstructured":"Diemo Schwarz and Baptiste Caramiaux . 2013 . Interactive sound texture synthesis through semi-automatic user annotations . In International Symposium on Computer Music Multidisciplinary Research. Springer, 372--392 . Diemo Schwarz and Baptiste Caramiaux. 2013. Interactive sound texture synthesis through semi-automatic user annotations. In International Symposium on Computer Music Multidisciplinary Research. Springer, 372--392."},{"key":"e_1_2_2_44_1","doi-asserted-by":"publisher","DOI":"10.1016\/j.ijhcs.2019.02.001"},{"key":"e_1_2_2_45_1","unstructured":"Paris Smaragdis Bhiksha Raj and Madhusudana Shashanka. 2006. A probabilistic latent variable model for acoustic modeling. (2006).  Paris Smaragdis Bhiksha Raj and Madhusudana Shashanka. 2006. A probabilistic latent variable model for acoustic modeling. (2006)."},{"key":"e_1_2_2_46_1","volume-title":"Proceedings of the 12th International Conference on Digital Audio Effects.","author":"Spiertz Martin","year":"2009","unstructured":"Martin Spiertz and Volker Gnann . 2009 . Source-filter based clustering for monaural blind source separation . In Proceedings of the 12th International Conference on Digital Audio Effects. Martin Spiertz and Volker Gnann. 2009. Source-filter based clustering for monaural blind source separation. In Proceedings of the 12th International Conference on Digital Audio Effects."},{"key":"e_1_2_2_47_1","volume-title":"Deep Audio Prior. ArXiv abs\/1912.10292","author":"Tian Yapeng","year":"2019","unstructured":"Yapeng Tian , Chenliang Xu , and Dingzeyu Li. 2019. Deep Audio Prior. ArXiv abs\/1912.10292 ( 2019 ). Yapeng Tian, Chenliang Xu, and Dingzeyu Li. 2019. Deep Audio Prior. ArXiv abs\/1912.10292 (2019)."},{"key":"e_1_2_2_48_1","doi-asserted-by":"publisher","DOI":"10.1109\/MMUL.2010.44"},{"volume-title":"Auditory Display","author":"Verron Charles","key":"e_1_2_2_49_1","unstructured":"Charles Verron , Mitsuko Aramaki , Richard Kronland-Martinet , and Gr\u00e9gory Pallone . 2009. Spatialized synthesis of noisy environmental sounds . In Auditory Display . Springer , 392--407. Charles Verron, Mitsuko Aramaki, Richard Kronland-Martinet, and Gr\u00e9gory Pallone. 2009. Spatialized synthesis of noisy environmental sounds. In Auditory Display. Springer, 392--407."},{"key":"e_1_2_2_50_1","doi-asserted-by":"publisher","DOI":"10.1145\/3197517.3201318"},{"key":"e_1_2_2_51_1","doi-asserted-by":"publisher","DOI":"10.1109\/ICME.2011.6011902"},{"key":"e_1_2_2_52_1","doi-asserted-by":"publisher","DOI":"10.1145\/3272127.3275090"},{"key":"e_1_2_2_53_1","doi-asserted-by":"crossref","first-page":"1","DOI":"10.1145\/3355089.3356566","article-title":"Acoustic texture rendering for extended sources in complex scenes","volume":"38","author":"Zhang Zechen","year":"2019","unstructured":"Zechen Zhang , Nikunj Raghuvanshi , John Snyder , and Steve Marschner . 2019 . Acoustic texture rendering for extended sources in complex scenes . ACM Transactions on Graphics (TOG) 38 , 6 (2019), 1 -- 9 . Zechen Zhang, Nikunj Raghuvanshi, John Snyder, and Steve Marschner. 2019. Acoustic texture rendering for extended sources in complex scenes. ACM Transactions on Graphics (TOG) 38, 6 (2019), 1--9.","journal-title":"ACM Transactions on Graphics (TOG)"},{"key":"e_1_2_2_54_1","doi-asserted-by":"publisher","DOI":"10.1145\/1576246.1531343"},{"key":"e_1_2_2_55_1","volume-title":"Proceedings of the 7th international conference on digital audio effects DAFX","volume":"4","author":"Zhu Xinglei","year":"2004","unstructured":"Xinglei Zhu and Lonce Wyse . 2004 . Sound texture modeling and time-frequency LPC . In Proceedings of the 7th international conference on digital audio effects DAFX , Vol. 4 . Xinglei Zhu and Lonce Wyse. 2004. Sound texture modeling and time-frequency LPC. In Proceedings of the 7th international conference on digital audio effects DAFX, Vol. 4."}],"container-title":["ACM Transactions on Graphics"],"original-title":[],"language":"en","link":[{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/3414685.3417875","content-type":"unspecified","content-version":"vor","intended-application":"text-mining"},{"URL":"https:\/\/dl.acm.org\/doi\/pdf\/10.1145\/3414685.3417875","content-type":"unspecified","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2025,6,17]],"date-time":"2025-06-17T22:03:17Z","timestamp":1750197797000},"score":1,"resource":{"primary":{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/3414685.3417875"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2020,11,27]]},"references-count":55,"journal-issue":{"issue":"6","published-print":{"date-parts":[[2020,12,31]]}},"alternative-id":["10.1145\/3414685.3417875"],"URL":"https:\/\/doi.org\/10.1145\/3414685.3417875","relation":{},"ISSN":["0730-0301","1557-7368"],"issn-type":[{"type":"print","value":"0730-0301"},{"type":"electronic","value":"1557-7368"}],"subject":[],"published":{"date-parts":[[2020,11,27]]},"assertion":[{"value":"2020-11-27","order":2,"name":"published","label":"Published","group":{"name":"publication_history","label":"Publication History"}}]}}