{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2026,7,31]],"date-time":"2026-07-31T08:58:13Z","timestamp":1785488293373,"version":"3.56.0"},"publisher-location":"New York, NY, USA","reference-count":33,"publisher":"ACM","license":[{"start":{"date-parts":[[2025,12,17]],"date-time":"2025-12-17T00:00:00Z","timestamp":1765929600000},"content-version":"vor","delay-in-days":0,"URL":"https:\/\/creativecommons.org\/licenses\/by-nc-nd\/4.0\/legalcode"}],"content-domain":{"domain":["dl.acm.org"],"crossmark-restriction":true},"short-container-title":[],"published-print":{"date-parts":[[2025,12,17]]},"DOI":"10.1145\/3774521.3774558","type":"proceedings-article","created":{"date-parts":[[2026,7,31]],"date-time":"2026-07-31T07:34:24Z","timestamp":1785483264000},"page":"1-8","update-policy":"https:\/\/doi.org\/10.1145\/crossmark-policy","source":"Crossref","is-referenced-by-count":0,"title":["Quality Aware Text-to-Image Synthesis with distortion maps as structural guides"],"prefix":"10.1145","author":[{"ORCID":"https:\/\/orcid.org\/0009-0002-3789-4223","authenticated-orcid":false,"given":"Kriti","family":"Khare","sequence":"first","affiliation":[{"name":"Indian Institute of Technology Mandi, Mandi, India"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"ORCID":"https:\/\/orcid.org\/0009-0001-4113-1600","authenticated-orcid":false,"given":"Parimala","family":"Kancharla","sequence":"additional","affiliation":[{"name":"Indian Institute of Technology, Mandi, Mandi, India"}],"role":[{"vocabulary":"crossref","role":"author"}]}],"member":"320","published-online":{"date-parts":[[2026,7,31]]},"reference":[{"key":"e_1_3_3_3_2_2","volume-title":"International Conference on Learning Representations","author":"Bi\u0144kowski Miko\u0142aj","year":"2018","unstructured":"Miko\u0142aj Bi\u0144kowski, Danica\u00a0J Sutherland, Michael Arbel, and Arthur Gretton. 2018. Demystifying MMD GANs. In International Conference on Learning Representations."},{"key":"e_1_3_3_3_3_2","doi-asserted-by":"publisher","unstructured":"Dominique Brunet Edward\u00a0R. Vrscay and Zhou Wang. 2012. On the Mathematical Properties of the Structural Similarity Index. IEEE Transactions on Image Processing 21 4 (2012) 1488\u20131499. 10.1109\/TIP.2011.2173206","DOI":"10.1109\/TIP.2011.2173206"},{"key":"e_1_3_3_3_4_2","doi-asserted-by":"publisher","DOI":"10.1007\/978-3-031-19836-66"},{"key":"e_1_3_3_3_5_2","doi-asserted-by":"publisher","unstructured":"Sathya Veera\u00a0Reddy Dendi Chander Dev Narayan Kothari and Sumohana\u00a0S. Channappayya. 2019. Generating Image Distortion Maps Using Convolutional Autoencoders With Application to No Reference Image Quality Assessment. IEEE Signal Processing Letters 26 1 (2019) 89\u201393. 10.1109\/LSP.2018.2879518","DOI":"10.1109\/LSP.2018.2879518"},{"key":"e_1_3_3_3_6_2","series-title":"(NIPS \u201921)","volume-title":"Proceedings of the 35th International Conference on Neural Information Processing Systems","author":"Dhariwal Prafulla","year":"2021","unstructured":"Prafulla Dhariwal and Alex Nichol. 2021. Diffusion models beat GANs on image synthesis. In Proceedings of the 35th International Conference on Neural Information Processing Systems(NIPS \u201921). Curran Associates Inc., Red Hook, NY, USA, Article 672, 15\u00a0pages."},{"key":"e_1_3_3_3_7_2","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR46437.2021.01268"},{"key":"e_1_3_3_3_8_2","doi-asserted-by":"publisher","DOI":"10.5555\/3295222.3295408"},{"key":"e_1_3_3_3_9_2","series-title":"(NIPS \u201920)","volume-title":"Proceedings of the 34th International Conference on Neural Information Processing Systems","author":"Ho Jonathan","year":"2020","unstructured":"Jonathan Ho, Ajay Jain, and Pieter Abbeel. 2020. Denoising diffusion probabilistic models. In Proceedings of the 34th International Conference on Neural Information Processing Systems (Vancouver, BC, Canada) (NIPS \u201920). Curran Associates Inc., Red Hook, NY, USA, Article 574, 12\u00a0pages."},{"key":"e_1_3_3_3_10_2","unstructured":"Jonathan Ho Tim Salimans Alexey Gritsenko William Chan Mohammad Norouzi and David\u00a0J Fleet. 2022. Video diffusion models. arXiv:https:\/\/arXiv.org\/abs\/2204.03458 (2022)."},{"key":"e_1_3_3_3_11_2","doi-asserted-by":"publisher","DOI":"10.1109\/ICCV51070.2023.00686"},{"key":"e_1_3_3_3_12_2","volume-title":"Quality aware generative adversarial networks","author":"Kancharla Parimala","year":"2019","unstructured":"Parimala Kancharla and Sumohana\u00a0S. Channappayya. 2019. Quality aware generative adversarial networks. Curran Associates Inc., Red Hook, NY, USA."},{"key":"e_1_3_3_3_13_2","volume-title":"International Conference on Learning Representations","author":"Karras Tero","year":"2018","unstructured":"Tero Karras, Timo Aila, Samuli Laine, and Jaakko Lehtinen. 2018. Progressive Growing of GANs for Improved Quality, Stability, and Variation. In International Conference on Learning Representations."},{"key":"e_1_3_3_3_14_2","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR.2019.00453"},{"key":"e_1_3_3_3_15_2","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR.2017.19"},{"key":"e_1_3_3_3_16_2","doi-asserted-by":"publisher","DOI":"10.1109\/FSKD.2018.8686914"},{"key":"e_1_3_3_3_17_2","doi-asserted-by":"publisher","DOI":"10.1007\/978-3-642-21735-77"},{"key":"e_1_3_3_3_18_2","volume-title":"The Twelfth International Conference on Learning Representations","author":"Podell Dustin","year":"2024","unstructured":"Dustin Podell, Zion English, Kyle Lacey, Andreas Blattmann, Tim Dockhorn, Jonas M\u00fcller, Joe Penna, and Robin Rombach. 2024. SDXL: Improving Latent Diffusion Models for High-Resolution Image Synthesis. In The Twelfth International Conference on Learning Representations. https:\/\/openreview.net\/forum?id=di52zR8xgf"},{"key":"e_1_3_3_3_19_2","series-title":"Proceedings of Machine Learning Research","first-page":"8748","volume-title":"Proceedings of the 38th International Conference on Machine Learning","volume":"139","author":"Radford Alec","year":"2021","unstructured":"Alec Radford, Jong\u00a0Wook Kim, Chris Hallacy, Aditya Ramesh, Gabriel Goh, Sandhini Agarwal, Girish Sastry, Amanda Askell, Pamela Mishkin, Jack Clark, Gretchen Krueger, and Ilya Sutskever. 2021. Learning Transferable Visual Models From Natural Language Supervision. In Proceedings of the 38th International Conference on Machine Learning(Proceedings of Machine Learning Research, Vol.\u00a0139), Marina Meila and Tong Zhang (Eds.). PMLR, 8748\u20138763. https:\/\/proceedings.mlr.press\/v139\/radford21a.html"},{"key":"e_1_3_3_3_20_2","series-title":"Proceedings of Machine Learning Research","first-page":"8821","volume-title":"Proceedings of the 38th International Conference on Machine Learning","volume":"139","author":"Ramesh Aditya","year":"2021","unstructured":"Aditya Ramesh, Mikhail Pavlov, Gabriel Goh, Scott Gray, Chelsea Voss, Alec Radford, Mark Chen, and Ilya Sutskever. 2021. Zero-Shot Text-to-Image Generation. In Proceedings of the 38th International Conference on Machine Learning(Proceedings of Machine Learning Research, Vol.\u00a0139), Marina Meila and Tong Zhang (Eds.). PMLR, 8821\u20138831. https:\/\/proceedings.mlr.press\/v139\/ramesh21a.html"},{"key":"e_1_3_3_3_21_2","series-title":"Proceedings of Machine Learning Research","first-page":"8821","volume-title":"Proceedings of the 38th International Conference on Machine Learning","volume":"139","author":"Ramesh Aditya","year":"2021","unstructured":"Aditya Ramesh, Mikhail Pavlov, Gabriel Goh, Scott Gray, Chelsea Voss, Alec Radford, Mark Chen, and Ilya Sutskever. 2021. Zero-Shot Text-to-Image Generation. In Proceedings of the 38th International Conference on Machine Learning(Proceedings of Machine Learning Research, Vol.\u00a0139), Marina Meila and Tong Zhang (Eds.). PMLR, 8821\u20138831. https:\/\/proceedings.mlr.press\/v139\/ramesh21a.html"},{"key":"e_1_3_3_3_22_2","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR52688.2022.01042"},{"key":"e_1_3_3_3_23_2","doi-asserted-by":"publisher","DOI":"10.1007\/978-3-319-24574-428"},{"key":"e_1_3_3_3_24_2","doi-asserted-by":"publisher","DOI":"10.1007\/978-0-387-30164-8_528"},{"key":"e_1_3_3_3_25_2","doi-asserted-by":"publisher","DOI":"10.1007\/978-3-031-73016-06"},{"key":"e_1_3_3_3_26_2","series-title":"Proceedings of Machine Learning Research","first-page":"2256","volume-title":"Proceedings of the 32nd International Conference on Machine Learning","volume":"37","author":"Sohl-Dickstein Jascha","year":"2015","unstructured":"Jascha Sohl-Dickstein, Eric Weiss, Niru Maheswaranathan, and Surya Ganguli. 2015. Deep Unsupervised Learning using Nonequilibrium Thermodynamics. In Proceedings of the 32nd International Conference on Machine Learning(Proceedings of Machine Learning Research, Vol.\u00a037), Francis Bach and David Blei (Eds.). PMLR, Lille, France, 2256\u20132265. https:\/\/proceedings.mlr.press\/v37\/sohl-dickstein15.html"},{"key":"e_1_3_3_3_27_2","doi-asserted-by":"publisher","DOI":"10.5555\/3327144.3327246"},{"key":"e_1_3_3_3_28_2","doi-asserted-by":"publisher","DOI":"10.5555\/3295222.3295349"},{"key":"e_1_3_3_3_29_2","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR.2017.404"},{"key":"e_1_3_3_3_30_2","doi-asserted-by":"publisher","DOI":"10.1109\/VCIP47243.2019.8965734"},{"key":"e_1_3_3_3_31_2","doi-asserted-by":"publisher","unstructured":"Zhou Wang A.C. Bovik H.R. Sheikh and E.P. Simoncelli. 2004. Image quality assessment: from error visibility to structural similarity. IEEE Transactions on Image Processing 13 4 (2004) 600\u2013612. 10.1109\/TIP.2003.819861","DOI":"10.1109\/TIP.2003.819861"},{"key":"e_1_3_3_3_32_2","doi-asserted-by":"publisher","unstructured":"Zhou Wang and Alan\u00a0C. Bovik. 2009. Mean squared error: Love it or leave it? A new look at Signal Fidelity Measures. IEEE Signal Processing Magazine 26 1 (2009) 98\u2013117. 10.1109\/MSP.2008.930649","DOI":"10.1109\/MSP.2008.930649"},{"key":"e_1_3_3_3_33_2","doi-asserted-by":"publisher","unstructured":"Cort\u00a0J. Willmott Steven\u00a0G. Ackleson Robert\u00a0E. Davis Johannes\u00a0J. Feddema Katherine\u00a0M. Klink David\u00a0R. Legates James O\u2019Donnell and Clinton\u00a0M. Rowe. 1985. Statistics for the evaluation and comparison of models. Journal of Geophysical Research: Oceans 90 C5 (Sept. 1985) 8995\u20139005. 10.1029\/jc090ic05p08995","DOI":"10.1029\/jc090ic05p08995"},{"key":"e_1_3_3_3_34_2","unstructured":"Fisher Yu Yinda Zhang Shuran Song Ari Seff and Jianxiong Xiao. 2015. LSUN: Construction of a Large-scale Image Dataset using Deep Learning with Humans in the Loop. arXiv preprint arXiv:https:\/\/arXiv.org\/abs\/1506.03365 (2015)."}],"event":{"name":"ICVGIP 2025: Indian Conference on Computer Vision, Graphics, and Image Processing","location":"Mandi Himachal Pradesh India","acronym":"ICVGIP 2025"},"container-title":["Proceedings of the Sixteen Indian Conference on Computer Vision, Graphics and Image Processing"],"original-title":[],"link":[{"URL":"https:\/\/dl.acm.org\/doi\/pdf\/10.1145\/3774521.3774558","content-type":"unspecified","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2026,7,31]],"date-time":"2026-07-31T08:08:28Z","timestamp":1785485308000},"score":1,"resource":{"primary":{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/3774521.3774558"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2025,12,17]]},"references-count":33,"alternative-id":["10.1145\/3774521.3774558","10.1145\/3774521"],"URL":"https:\/\/doi.org\/10.1145\/3774521.3774558","relation":{},"subject":[],"published":{"date-parts":[[2025,12,17]]},"assertion":[{"value":"2026-07-31","order":3,"name":"published","label":"Published","group":{"name":"publication_history","label":"Publication History"}}]}}