{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2025,8,24]],"date-time":"2025-08-24T00:02:35Z","timestamp":1755993755246,"version":"3.44.0"},"reference-count":25,"publisher":"Association for Computing Machinery (ACM)","issue":"3","license":[{"start":{"date-parts":[[2024,8,9]],"date-time":"2024-08-09T00:00:00Z","timestamp":1723161600000},"content-version":"vor","delay-in-days":0,"URL":"https:\/\/www.acm.org\/publications\/policies\/copyright_policy#Background"}],"content-domain":{"domain":["dl.acm.org"],"crossmark-restriction":true},"short-container-title":["Proc. ACM Comput. Graph. Interact. Tech."],"published-print":{"date-parts":[[2024,8,9]]},"abstract":"<jats:p>As technology advances from simple 2D designs to intricate 3D environments, the demand for high-quality visuals in video games and interactive media necessitates robust image quality assessment (IQA) techniques. Traditional methods like PSNR and SSIM, reliant on reference images, struggle with the unique challenges of 3D rendered content, highlighting the need for specialized non-reference IQA approaches. This paper introduces a novel multi-task learning architecture that corrects and predicts aliasing artifacts simultaneously, enhancing predictive accuracy without reference images. It also incorporates temporal information to improve visual coherence and smoothness. An automated labeling pipeline developed using Unity ensures a stable and unbiased dataset for model training and evaluation. Our experiments demonstrate that this approach reliably detects aliasing across various complexities, achieving state-of-the-art performance. By addressing specific challenges in rendered image assessment and leveraging innovative learning techniques, our work advances IQA for video games and simulations, ensuring high visual quality.<\/jats:p>","DOI":"10.1145\/3675379","type":"journal-article","created":{"date-parts":[[2024,8,9]],"date-time":"2024-08-09T15:53:18Z","timestamp":1723218798000},"page":"1-12","update-policy":"https:\/\/doi.org\/10.1145\/crossmark-policy","source":"Crossref","is-referenced-by-count":0,"title":["Aliasing Detection in Rendered Images via a Multi-Task Learning"],"prefix":"10.1145","volume":"7","author":[{"ORCID":"https:\/\/orcid.org\/0009-0005-9388-6123","authenticated-orcid":false,"given":"Shu-Ho","family":"Fan","sequence":"first","affiliation":[{"name":"National Tsing-Hua University, Taiwan"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"ORCID":"https:\/\/orcid.org\/0000-0002-3007-7274","authenticated-orcid":false,"given":"Kai-Wen","family":"Hsiao","sequence":"additional","affiliation":[{"name":"National Tsing-Hua University, Taiwan"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"ORCID":"https:\/\/orcid.org\/0000-0002-0973-1892","authenticated-orcid":false,"given":"Kai Yi","family":"Tan","sequence":"additional","affiliation":[{"name":"Tunghai University, Taiwan"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"ORCID":"https:\/\/orcid.org\/0000-0001-6608-7178","authenticated-orcid":false,"given":"Chih-Yuan","family":"Yao","sequence":"additional","affiliation":[{"name":"National Taiwan University of Science and Technology, Taiwan"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"ORCID":"https:\/\/orcid.org\/0000-0001-7153-4411","authenticated-orcid":false,"given":"Hung-Kuo","family":"Chu","sequence":"additional","affiliation":[{"name":"National Tsing-Hua University, Taiwan"}],"role":[{"role":"author","vocabulary":"crossref"}]}],"member":"320","published-online":{"date-parts":[[2024,8,9]]},"reference":[{"key":"e_1_2_2_1_1","unstructured":"AMD. 2023. FidelityFX Super Resolution 2. https:\/\/gpuopen.com\/fidelityfx-superresolution-2\/."},{"key":"e_1_2_2_2_1","volume-title":"Proc. ACM Comput. Graph. Interact. Tech. 3 (2020","author":"Andersson Pontus","year":"2064","unstructured":"Pontus Andersson, Jim Nilsson, Tomas Akenine-M\u00f6ller, Magnus Oskarsson, Karl Johan \u00c5str\u00f6m, and Mark D. Fairchild. 2020. FLIP: A Difference Evaluator for Alternating Images. Proc. ACM Comput. Graph. Interact. Tech. 3 (2020), 15:1--15:23. https:\/\/api.semanticscholar.org\/CorpusID:220643528"},{"key":"e_1_2_2_3_1","doi-asserted-by":"publisher","DOI":"10.1109\/ICCV.2019.00325"},{"key":"e_1_2_2_4_1","doi-asserted-by":"publisher","DOI":"10.1145\/3072959.3073708"},{"key":"e_1_2_2_5_1","doi-asserted-by":"publisher","DOI":"10.1109\/TIP.2016.2598681"},{"key":"e_1_2_2_6_1","volume-title":"ECCV Workshops. https:\/\/api.semanticscholar.org\/CorpusID:252519482","author":"Conde Marcos V.","year":"2022","unstructured":"Marcos V. Conde, Ui-Jin Choi, Maxime Burchi, and Radu Timofte. 2022. Swin2SR: SwinV2 Transformer for Compressed Image Super-Resolution and Restoration. In ECCV Workshops. https:\/\/api.semanticscholar.org\/CorpusID:252519482"},{"key":"e_1_2_2_7_1","unstructured":"Killian Herveau Max Piochowiak and Carsten Dachsbacher. 2023. Minimal Convolutional Neural Networks for Temporal Anti Aliasing. https:\/\/api.semanticscholar.org\/CorpusID:259305872"},{"key":"e_1_2_2_8_1","doi-asserted-by":"crossref","unstructured":"Jiaxi Jiang Kai Zhang and Radu Timofte. 2021. Towards Flexible Blind JPEG Artifacts Removal. arXiv:2109.14573 [eess.IV]","DOI":"10.1109\/ICCV48922.2021.00495"},{"key":"e_1_2_2_9_1","volume-title":"Unity: A General Platform for Intelligent Agents. arXiv:1809.02627 [cs.LG]","author":"Juliani Arthur","year":"2020","unstructured":"Arthur Juliani, Vincent-Pierre Berges, Ervin Teng, Andrew Cohen, Jonathan Harper, Chris Elion, Chris Goy, Yuan Gao, Hunter Henry, Marwan Mattar, and Danny Lange. 2020. Unity: A General Platform for Intelligent Agents. arXiv:1809.02627 [cs.LG]"},{"key":"e_1_2_2_10_1","volume-title":"SwinIR: Image Restoration Using Swin Transformer. 2021 IEEE\/CVF International Conference on Computer Vision Workshops (ICCVW) (2021)","author":"Liang Jingyun","year":"2021","unstructured":"Jingyun Liang, Jie Cao, Guolei Sun, K. Zhang, Luc Van Gool, and Radu Timofte. 2021. SwinIR: Image Restoration Using Swin Transformer. 2021 IEEE\/CVF International Conference on Computer Vision Workshops (ICCVW) (2021), 1833--1844. https:\/\/api.semanticscholar.org\/CorpusID:237266491"},{"key":"e_1_2_2_11_1","volume-title":"Chen Change Loy, and Thomas S. Huang","author":"Liu Ding","year":"2018","unstructured":"Ding Liu, Bihan Wen, Yuchen Fan, Chen Change Loy, and Thomas S. Huang. 2018a. Non-Local Recurrent Network for Image Restoration. In Neural Information Processing Systems. https:\/\/api.semanticscholar.org\/CorpusID:47007607"},{"key":"e_1_2_2_12_1","volume-title":"Multi-level Wavelet-CNN for Image Restoration. 2018 IEEE\/CVF Conference on Computer Vision and Pattern Recognition Workshops (CVPRW) (2018","author":"Liu Pengju","year":"2018","unstructured":"Pengju Liu, Hongzhi Zhang, K. Zhang, Liang Lin, and Wangmeng Zuo. 2018b. Multi-level Wavelet-CNN for Image Restoration. 2018 IEEE\/CVF Conference on Computer Vision and Pattern Recognition Workshops (CVPRW) (2018), 886--88609. https:\/\/api.semanticscholar.org\/CorpusID:29151865"},{"key":"e_1_2_2_13_1","doi-asserted-by":"publisher","DOI":"10.1145\/3450626.3459831"},{"key":"e_1_2_2_14_1","doi-asserted-by":"publisher","DOI":"10.1145\/1964921.1964935"},{"key":"e_1_2_2_15_1","volume-title":"Real-time Monte Carlo Denoising with the Neural Bilateral Grid. In Eurographics Symposium on Rendering. https:\/\/api.semanticscholar.org\/CorpusID:220284605","author":"Meng Xiaoxu","year":"2020","unstructured":"Xiaoxu Meng, Quan Zheng, Amitabh Varshney, Gurprit Singh, and Matthias Zwicker. 2020. Real-time Monte Carlo Denoising with the Neural Bilateral Grid. In Eurographics Symposium on Rendering. https:\/\/api.semanticscholar.org\/CorpusID:220284605"},{"key":"e_1_2_2_16_1","doi-asserted-by":"publisher","DOI":"10.1145\/3231578.3231580"},{"key":"e_1_2_2_17_1","volume-title":"Very Deep Convolutional Networks for Large-Scale Image Recognition. CoRR abs\/1409.1556","author":"Simonyan Karen","year":"2014","unstructured":"Karen Simonyan and Andrew Zisserman. 2014. Very Deep Convolutional Networks for Large-Scale Image Recognition. CoRR abs\/1409.1556 (2014). https:\/\/api.semanticscholar.org\/CorpusID:14124313"},{"key":"e_1_2_2_18_1","doi-asserted-by":"publisher","DOI":"10.1145\/3414685.3417786"},{"key":"e_1_2_2_19_1","volume-title":"Classifier Guided Temporal Supersampling for Real-time Rendering. Computer Graphics Forum 41","author":"Vouga Etienne","year":"2022","unstructured":"Etienne Vouga, Christopher Wojtan, Yu-Xiao Guo, Guojun Chen, Yue Dong, and Xin Tong. 2022. Classifier Guided Temporal Supersampling for Real-time Rendering. Computer Graphics Forum 41 (2022). https:\/\/api.semanticscholar.org\/CorpusID:254249752"},{"key":"e_1_2_2_20_1","doi-asserted-by":"publisher","DOI":"10.1109\/TIP.2003.819861"},{"key":"e_1_2_2_21_1","volume-title":"Uformer: A General U-Shaped Transformer for Image Restoration. 2022 IEEE\/CVF Conference on Computer Vision and Pattern Recognition (CVPR)","author":"Wang Zhendong","year":"2021","unstructured":"Zhendong Wang, Xiaodong Cun, Jianmin Bao, and Jianzhuang Liu. 2021. Uformer: A General U-Shaped Transformer for Image Restoration. 2022 IEEE\/CVF Conference on Computer Vision and Pattern Recognition (CVPR) (2021), 17662--17672. https:\/\/api.semanticscholar.org\/CorpusID:235358213"},{"key":"e_1_2_2_22_1","doi-asserted-by":"publisher","DOI":"10.1145\/3386569.3392376"},{"key":"e_1_2_2_23_1","volume-title":"Variational Denoising Network: Toward Blind Noise Modeling and Removal. ArXiv abs\/1908.11314","author":"Yue Zongsheng","year":"2019","unstructured":"Zongsheng Yue, Hongwei Yong, Qian Zhao, Lei Zhang, and Deyu Meng. 2019. Variational Denoising Network: Toward Blind Noise Modeling and Removal. ArXiv abs\/1908.11314 (2019). https:\/\/api.semanticscholar.org\/CorpusID:201667906"},{"key":"e_1_2_2_24_1","volume-title":"Restormer: Efficient Transformer for High-Resolution Image Restoration. 2022 IEEE\/CVF Conference on Computer Vision and Pattern Recognition (CVPR)","author":"Zamir Syed Waqas","year":"2021","unstructured":"Syed Waqas Zamir, Aditya Arora, Salman Hameed Khan, Munawar Hayat, Fahad Shahbaz Khan, and Ming-Hsuan Yang. 2021. Restormer: Efficient Transformer for High-Resolution Image Restoration. 2022 IEEE\/CVF Conference on Computer Vision and Pattern Recognition (CVPR) (2021), 5718--5729. https:\/\/api.semanticscholar.org\/CorpusID:244346144"},{"key":"e_1_2_2_25_1","doi-asserted-by":"crossref","unstructured":"Richard Zhang Phillip Isola Alexei A Efros Eli Shechtman and Oliver Wang. 2018. The Unreasonable Effectiveness of Deep Features as a Perceptual Metric. In CVPR.","DOI":"10.1109\/CVPR.2018.00068"}],"container-title":["Proceedings of the ACM on Computer Graphics and Interactive Techniques"],"original-title":[],"language":"en","link":[{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/3675379","content-type":"unspecified","content-version":"vor","intended-application":"text-mining"},{"URL":"https:\/\/dl.acm.org\/doi\/pdf\/10.1145\/3675379","content-type":"unspecified","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2025,8,23]],"date-time":"2025-08-23T02:03:33Z","timestamp":1755914613000},"score":1,"resource":{"primary":{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/3675379"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2024,8,9]]},"references-count":25,"journal-issue":{"issue":"3","published-print":{"date-parts":[[2024,8,9]]}},"alternative-id":["10.1145\/3675379"],"URL":"https:\/\/doi.org\/10.1145\/3675379","relation":{},"ISSN":["2577-6193"],"issn-type":[{"type":"electronic","value":"2577-6193"}],"subject":[],"published":{"date-parts":[[2024,8,9]]},"assertion":[{"value":"2024-08-09","order":3,"name":"published","label":"Published","group":{"name":"publication_history","label":"Publication History"}}]}}