{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2026,7,17]],"date-time":"2026-07-17T06:09:31Z","timestamp":1784268571944,"version":"3.55.0"},"reference-count":54,"publisher":"Association for Computing Machinery (ACM)","issue":"4","license":[{"start":{"date-parts":[[2024,7,19]],"date-time":"2024-07-19T00:00:00Z","timestamp":1721347200000},"content-version":"vor","delay-in-days":0,"URL":"https:\/\/creativecommons.org\/licenses\/by\/4.0\/"}],"funder":[{"DOI":"10.13039\/100000001","name":"NSF","doi-asserted-by":"publisher","award":["2105806"],"award-info":[{"award-number":["2105806"]}],"id":[{"id":"10.13039\/100000001","id-type":"DOI","asserted-by":"publisher"}]}],"content-domain":{"domain":["dl.acm.org"],"crossmark-restriction":true},"short-container-title":["ACM Trans. Graph."],"published-print":{"date-parts":[[2024,7,19]]},"abstract":"<jats:p>Metropolis Light Transport (MLT) is a global illumination algorithm that is well-known for rendering challenging scenes with intricate light paths. However, MLT methods tend to produce unpredictable correlation artifacts in images, which can introduce visual inconsistencies for animation rendering. This drawback also makes it challenging to denoise MLT renderings while maintaining temporal stability. We tackle this issue with modern learning-based methods and build a sequence denoiser combining the recurrent connections with the cutting-edge vision transformer architecture. We demonstrate that our sophisticated denoiser can consistently improve the quality and temporal stability of MLT renderings with difficult light paths. Our method is efficient and scalable for complex scene renderings that require high sample counts.<\/jats:p>","DOI":"10.1145\/3658218","type":"journal-article","created":{"date-parts":[[2024,7,19]],"date-time":"2024-07-19T14:47:57Z","timestamp":1721400477000},"page":"1-14","update-policy":"https:\/\/doi.org\/10.1145\/crossmark-policy","source":"Crossref","is-referenced-by-count":5,"title":["Temporally Stable Metropolis Light Transport Denoising using Recurrent Transformer Blocks"],"prefix":"10.1145","volume":"43","author":[{"ORCID":"https:\/\/orcid.org\/0009-0004-4544-9378","authenticated-orcid":false,"given":"Chuhao","family":"Chen","sequence":"first","affiliation":[{"name":"University of California San Diego, San Diego, United States of America"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"ORCID":"https:\/\/orcid.org\/0009-0005-1575-6112","authenticated-orcid":false,"given":"Yuze","family":"He","sequence":"additional","affiliation":[{"name":"Tsinghua University, Beijing, China"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"ORCID":"https:\/\/orcid.org\/0000-0001-5443-470X","authenticated-orcid":false,"given":"Tzu-Mao","family":"Li","sequence":"additional","affiliation":[{"name":"University of California San Diego, San Diego, United States of America"}],"role":[{"vocabulary":"crossref","role":"author"}]}],"member":"320","published-online":{"date-parts":[[2024,7,19]]},"reference":[{"key":"e_1_2_2_1_1","doi-asserted-by":"publisher","DOI":"10.1016\/S0097-8493(00)00131-X"},{"key":"e_1_2_2_2_1","doi-asserted-by":"publisher","DOI":"10.1145\/3414685.3417847"},{"key":"e_1_2_2_3_1","volume-title":"Input-Dependent Uncorrelated Weighting for Monte Carlo Denoising. In SIGGRAPH Asia Conference Proceedings. 1--10","author":"Back Jonghee","year":"2023","unstructured":"Jonghee Back, Binh-Son Hua, Toshiya Hachisuka, and Bochang Moon. 2023. Input-Dependent Uncorrelated Weighting for Monte Carlo Denoising. In SIGGRAPH Asia Conference Proceedings. 1--10."},{"key":"e_1_2_2_4_1","doi-asserted-by":"publisher","DOI":"10.1145\/3072959.3073708"},{"key":"e_1_2_2_5_1","doi-asserted-by":"crossref","unstructured":"Martin Balint Krzysztof Wolski Karol Myszkowski Hans-Peter Seidel and Rafa\u0142 Mantiuk. 2023. Neural partitioning pyramids for denoising Monte Carlo renderings. 1--11.","DOI":"10.1145\/3588432.3591562"},{"key":"e_1_2_2_6_1","article-title":"Ensemble Metropolis Light Transport","volume":"41","author":"Bashford-Rogers Thomas","year":"2021","unstructured":"Thomas Bashford-Rogers, Lu\u00eds Paulo Santos, Demetris Marnerides, and Kurt Debattista. 2021. Ensemble Metropolis Light Transport. ACM Trans. Graph. 41, 1, Article 5 (2021), 15 pages.","journal-title":"ACM Trans. Graph."},{"key":"e_1_2_2_7_1","unstructured":"Benedikt Bitterli. 2016. VeachAjar. https:\/\/benedikt-bitterli.me\/resources\/."},{"key":"e_1_2_2_8_1","article-title":"Selectively Metropolised Monte Carlo Light Transport Simulation","volume":"38","author":"Bitterli Benedikt","year":"2019","unstructured":"Benedikt Bitterli and Wojciech Jarosz. 2019. Selectively Metropolised Monte Carlo Light Transport Simulation. ACM Trans. Graph. (Proc. SIGGRAPH Asia) 38, 6, Article 153 (2019), 10 pages.","journal-title":"ACM Trans. Graph. (Proc. SIGGRAPH Asia)"},{"key":"e_1_2_2_9_1","article-title":"Interactive Reconstruction of Monte Carlo Image Sequences Using a Recurrent Denoising Autoencoder","volume":"36","author":"Alla Chaitanya Chakravarty R.","year":"2017","unstructured":"Chakravarty R. Alla Chaitanya, Anton S. Kaplanyan, Christoph Schied, Marco Salvi, Aaron Lefohn, Derek Nowrouzezahrai, and Timo Aila. 2017. Interactive Reconstruction of Monte Carlo Image Sequences Using a Recurrent Denoising Autoencoder. ACM Trans. Graph. (Proc. SIGGRAPH) 36, 4, Article 98 (2017), 12 pages.","journal-title":"ACM Trans. Graph. (Proc. SIGGRAPH)"},{"key":"e_1_2_2_10_1","doi-asserted-by":"publisher","DOI":"10.1145\/1073204.1073330"},{"key":"e_1_2_2_11_1","unstructured":"Markus Ebke. 2021. LightSheet. https:\/\/benedikt-bitterli.me\/resources\/."},{"key":"e_1_2_2_12_1","doi-asserted-by":"publisher","DOI":"10.1111\/cgf.14338"},{"key":"e_1_2_2_13_1","doi-asserted-by":"publisher","DOI":"10.1145\/3306346.3322954"},{"key":"e_1_2_2_14_1","doi-asserted-by":"publisher","DOI":"10.1111\/cgf.13935"},{"key":"e_1_2_2_15_1","doi-asserted-by":"publisher","DOI":"10.1145\/2601097.2601138"},{"key":"e_1_2_2_16_1","doi-asserted-by":"publisher","DOI":"10.1111\/cgf.13919"},{"key":"e_1_2_2_17_1","unstructured":"Kaiming He Xiangyu Zhang Shaoqing Ren and Jian Sun. 2016. Deep residual learning for image recognition. In Computer Vision and Pattern Recognition. 770--778."},{"key":"e_1_2_2_18_1","doi-asserted-by":"publisher","DOI":"10.1145\/3450626.3459793"},{"key":"e_1_2_2_19_1","unstructured":"Wenzel Jakob. 2010. Mitsuba renderer. http:\/\/www.mitsuba-renderer.org."},{"key":"e_1_2_2_20_1","doi-asserted-by":"publisher","DOI":"10.1145\/2185520.2185554"},{"key":"e_1_2_2_21_1","doi-asserted-by":"publisher","DOI":"10.1145\/2766977"},{"key":"e_1_2_2_22_1","doi-asserted-by":"publisher","DOI":"10.1111\/1467-8659.t01-1-00703"},{"key":"e_1_2_2_23_1","volume-title":"Adam: A method for stochastic optimization.","author":"Kingma Diederick P","year":"2015","unstructured":"Diederick P Kingma and Jimmy Ba. 2015. Adam: A method for stochastic optimization. (2015)."},{"key":"e_1_2_2_24_1","doi-asserted-by":"publisher","DOI":"10.1111\/j.1467-8659.2009.01540.x"},{"key":"e_1_2_2_25_1","doi-asserted-by":"publisher","DOI":"10.1145\/3269978"},{"key":"e_1_2_2_27_1","unstructured":"Zhengqin Li Ting-Wei Yu Shen Sang Sarah Wang Meng Song Yuhan Liu Yu-Ying Yeh Rui Zhu Nitesh Gundavarapu Jia Shi et al. 2021. OpenRooms: An open framework for photorealistic indoor scene datasets. In Computer Vision and Pattern Recognition. 7190--7199."},{"key":"e_1_2_2_28_1","volume-title":"VRT: A video restoration transformer. arXiv preprint arXiv:2201.12288","author":"Liang Jingyun","year":"2022","unstructured":"Jingyun Liang, Jiezhang Cao, Yuchen Fan, Kai Zhang, Rakesh Ranjan, Yawei Li, Radu Timofte, and Luc Van Gool. 2022a. VRT: A video restoration transformer. arXiv preprint arXiv:2201.12288 (2022)."},{"key":"e_1_2_2_29_1","doi-asserted-by":"publisher","DOI":"10.1109\/ICCVW54120.2021.00210"},{"key":"e_1_2_2_30_1","first-page":"378","article-title":"Recurrent video restoration transformer with guided deformable attention","volume":"35","author":"Liang Jingyun","year":"2022","unstructured":"Jingyun Liang, Yuchen Fan, Xiaoyu Xiang, Rakesh Ranjan, Eddy Ilg, Simon Green, Jiezhang Cao, Kai Zhang, Radu Timofte, and Luc V Gool. 2022b. Recurrent video restoration transformer with guided deformable attention. Advances in Neural Information Processing Systems 35 (2022), 378--393.","journal-title":"Advances in Neural Information Processing Systems"},{"key":"e_1_2_2_31_1","doi-asserted-by":"publisher","DOI":"10.1109\/ICCV48922.2021.00986"},{"key":"e_1_2_2_32_1","volume-title":"A ConvNet for the","author":"Liu Zhuang","year":"2020","unstructured":"Zhuang Liu, Hanzi Mao, Chao-Yuan Wu, Christoph Feichtenhofer, Trevor Darrell, and Saining Xie. 2022. A ConvNet for the 2020s. In Computer Vision and Pattern Recognition. 11976--11986."},{"key":"e_1_2_2_33_1","doi-asserted-by":"publisher","DOI":"10.1145\/3386569.3392382"},{"key":"e_1_2_2_34_1","doi-asserted-by":"publisher","DOI":"10.1145\/3450626.3459831"},{"key":"e_1_2_2_35_1","doi-asserted-by":"publisher","DOI":"10.1145\/2366145.2366182"},{"key":"e_1_2_2_36_1","volume-title":"Real-time Monte Carlo Denoising with the Neural Bilateral Grid. In Eurographics Symposium on Rendering - DL-only Track. 13--24","author":"Meng Xiaoxu","year":"2020","unstructured":"Xiaoxu Meng, Quan Zheng, Amitabh Varshney, Gurprit Singh, and Matthias Zwicker. 2020. Real-time Monte Carlo Denoising with the Neural Bilateral Grid. In Eurographics Symposium on Rendering - DL-only Track. 13--24."},{"key":"e_1_2_2_37_1","volume-title":"PyTorch: An Imperative Style","author":"Paszke Adam","unstructured":"Adam Paszke, Sam Gross, Francisco Massa, Adam Lerer, James Bradbury, Gregory Chanan, Trevor Killeen, Zeming Lin, Natalia Gimelshein, Luca Antiga, Alban Desmaison, Andreas Kopf, Edward Yang, Zachary DeVito, Martin Raison, Alykhan Tejani, Sasank Chilamkurthy, Benoit Steiner, Lu Fang, Junjie Bai, and Soumith Chintala. 2019. PyTorch: An Imperative Style, High-Performance Deep Learning Library. In Advances in Neural Information Processing Systems. 8024--8035."},{"key":"e_1_2_2_38_1","doi-asserted-by":"publisher","DOI":"10.1007\/978-3-319-24574-4_28"},{"key":"e_1_2_2_39_1","volume-title":"John Burgess, Shiqiu Liu, Carsten Dachsbacher, Aaron Lefohn, and Marco Salvi.","author":"Schied Christoph","year":"2017","unstructured":"Christoph Schied, Anton Kaplanyan, Chris Wyman, Anjul Patney, Chakravarty R Alla Chaitanya, John Burgess, Shiqiu Liu, Carsten Dachsbacher, Aaron Lefohn, and Marco Salvi. 2017. Spatiotemporal variance-guided filtering: real-time reconstruction for path-traced global illumination. In High Performance Graphics. 1--12."},{"key":"e_1_2_2_40_1","first-page":"1","article-title":"Temporally stable real-time joint neural denoising and supersampling. Proceedings of the ACM on Computer Graphics and Interactive Techniques (Proc","volume":"5","author":"Thomas Manu Mathew","year":"2022","unstructured":"Manu Mathew Thomas, Gabor Liktor, Christoph Peters, Sungye Kim, Karthik Vaidyanathan, and Angus G Forbes. 2022. Temporally stable real-time joint neural denoising and supersampling. Proceedings of the ACM on Computer Graphics and Interactive Techniques (Proc. HPG) 5, 3 (2022), 1--22.","journal-title":"HPG)"},{"key":"e_1_2_2_41_1","volume-title":"International Conference on Machine Learning. 10347--10357","author":"Touvron Hugo","year":"2021","unstructured":"Hugo Touvron, Matthiegu Cord, Matthijs Douze, Francisco Massa, Alexandre Sablayrolles, and Herv\u00e9 J\u00e9gou. 2021. Training data-efficient image transformers & distillation through attention. In International Conference on Machine Learning. 10347--10357."},{"key":"e_1_2_2_42_1","volume-title":"Eurographics Symposium on Rendering - Experimental Ideas & Implementations. 55--63","author":"de Woestijne Joran Van","year":"2017","unstructured":"Joran Van de Woestijne, Roald Frederickx, Niels Billen, and Philip Dutr\u00e9. 2017. Temporal coherence for Metropolis light transport. In Eurographics Symposium on Rendering - Experimental Ideas & Implementations. 55--63."},{"key":"e_1_2_2_43_1","volume-title":"Attention is all you need. Advances in neural information processing systems 30","author":"Vaswani Ashish","year":"2017","unstructured":"Ashish Vaswani, Noam Shazeer, Niki Parmar, Jakob Uszkoreit, Llion Jones, Aidan N Gomez, \u0141ukasz Kaiser, and Illia Polosukhin. 2017. Attention is all you need. Advances in neural information processing systems 30 (2017)."},{"key":"e_1_2_2_44_1","volume-title":"Robust Monte Carlo methods for light transport simulation","author":"Veach Eric","unstructured":"Eric Veach. 1998. Robust Monte Carlo methods for light transport simulation. Stanford University."},{"key":"e_1_2_2_45_1","doi-asserted-by":"crossref","unstructured":"Eric Veach and Leonidas J Guibas. 1997. Metropolis light transport. In SIGGRAPH. 65--76.","DOI":"10.1145\/258734.258775"},{"key":"e_1_2_2_46_1","doi-asserted-by":"publisher","DOI":"10.1145\/3197517.3201388"},{"key":"e_1_2_2_47_1","volume-title":"Uformer: A general U-shaped transformer for image restoration. In Computer Vision and Pattern Recognition. 17683--17693.","author":"Wang Zhendong","year":"2022","unstructured":"Zhendong Wang, Xiaodong Cun, Jianmin Bao, Wengang Zhou, Jianzhuang Liu, and Houqiang Li. 2022. Uformer: A general U-shaped transformer for image restoration. In Computer Vision and Pattern Recognition. 17683--17693."},{"key":"e_1_2_2_48_1","doi-asserted-by":"publisher","DOI":"10.1111\/cgf.13232"},{"key":"e_1_2_2_49_1","doi-asserted-by":"publisher","DOI":"10.1145\/3386569.3392376"},{"key":"e_1_2_2_50_1","doi-asserted-by":"publisher","DOI":"10.1145\/2816814"},{"key":"e_1_2_2_51_1","doi-asserted-by":"publisher","DOI":"10.1145\/3478513.3480565"},{"key":"e_1_2_2_52_1","volume-title":"Fahad Shahbaz Khan, and Ming-Hsuan Yang","author":"Zamir Syed Waqas","year":"2022","unstructured":"Syed Waqas Zamir, Aditya Arora, Salman Khan, Munawar Hayat, Fahad Shahbaz Khan, and Ming-Hsuan Yang. 2022. Restormer: Efficient transformer for high-resolution image restoration. In Computer Vision and Pattern Recognition. 5728--5739."},{"key":"e_1_2_2_53_1","doi-asserted-by":"publisher","DOI":"10.5555\/2858834.2858848"},{"key":"e_1_2_2_54_1","doi-asserted-by":"publisher","DOI":"10.1145\/3414685.3417856"},{"key":"e_1_2_2_55_1","doi-asserted-by":"publisher","DOI":"10.5555\/2816723.2816781"}],"container-title":["ACM Transactions on Graphics"],"original-title":[],"language":"en","link":[{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/3658218","content-type":"unspecified","content-version":"vor","intended-application":"text-mining"},{"URL":"https:\/\/dl.acm.org\/doi\/pdf\/10.1145\/3658218","content-type":"application\/pdf","content-version":"vor","intended-application":"syndication"},{"URL":"https:\/\/dl.acm.org\/doi\/pdf\/10.1145\/3658218","content-type":"unspecified","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2025,6,19]],"date-time":"2025-06-19T00:04:16Z","timestamp":1750291456000},"score":1,"resource":{"primary":{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/3658218"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2024,7,19]]},"references-count":54,"journal-issue":{"issue":"4","published-print":{"date-parts":[[2024,7,19]]}},"alternative-id":["10.1145\/3658218"],"URL":"https:\/\/doi.org\/10.1145\/3658218","relation":{},"ISSN":["0730-0301","1557-7368"],"issn-type":[{"value":"0730-0301","type":"print"},{"value":"1557-7368","type":"electronic"}],"subject":[],"published":{"date-parts":[[2024,7,19]]},"assertion":[{"value":"2024-07-19","order":3,"name":"published","label":"Published","group":{"name":"publication_history","label":"Publication History"}}]}}