{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2025,6,18]],"date-time":"2025-06-18T04:08:52Z","timestamp":1750219732077,"version":"3.41.0"},"publisher-location":"New York, NY, USA","reference-count":32,"publisher":"ACM","license":[{"start":{"date-parts":[[2023,10,29]],"date-time":"2023-10-29T00:00:00Z","timestamp":1698537600000},"content-version":"vor","delay-in-days":0,"URL":"https:\/\/www.acm.org\/publications\/policies\/copyright_policy#Background"}],"funder":[{"DOI":"10.13039\/501100001691","name":"Japan Society for the Promotion of Science","doi-asserted-by":"publisher","award":["22H03612"],"award-info":[{"award-number":["22H03612"]}],"id":[{"id":"10.13039\/501100001691","id-type":"DOI","asserted-by":"publisher"}]}],"content-domain":{"domain":["dl.acm.org"],"crossmark-restriction":true},"short-container-title":[],"published-print":{"date-parts":[[2023,10,29]]},"DOI":"10.1145\/3607541.3616818","type":"proceedings-article","created":{"date-parts":[[2023,10,17]],"date-time":"2023-10-17T22:10:31Z","timestamp":1697580631000},"page":"115-125","update-policy":"https:\/\/doi.org\/10.1145\/crossmark-policy","source":"Crossref","is-referenced-by-count":2,"title":["Nonword-to-Image Generation Considering Perceptual Association of Phonetically Similar Words"],"prefix":"10.1145","author":[{"ORCID":"https:\/\/orcid.org\/0000-0003-2453-4560","authenticated-orcid":false,"given":"Chihaya","family":"Matsuhira","sequence":"first","affiliation":[{"name":"Nagoya University, Nagoya, Japan"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"ORCID":"https:\/\/orcid.org\/0000-0002-9193-5973","authenticated-orcid":false,"given":"Marc A.","family":"Kastner","sequence":"additional","affiliation":[{"name":"Kyoto University, Kyoto, Japan"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"ORCID":"https:\/\/orcid.org\/0000-0002-3041-4330","authenticated-orcid":false,"given":"Takahiro","family":"Komamizu","sequence":"additional","affiliation":[{"name":"Nagoya University, Nagoya, Japan"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"ORCID":"https:\/\/orcid.org\/0000-0001-6290-9680","authenticated-orcid":false,"given":"Takatsugu","family":"Hirayama","sequence":"additional","affiliation":[{"name":"University of Human Environments, Okazaki, Japan"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"ORCID":"https:\/\/orcid.org\/0000-0001-6040-4988","authenticated-orcid":false,"given":"Keisuke","family":"Doman","sequence":"additional","affiliation":[{"name":"Chukyo University, Toyota, Japan"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"ORCID":"https:\/\/orcid.org\/0000-0003-3942-9296","authenticated-orcid":false,"given":"Ichiro","family":"Ide","sequence":"additional","affiliation":[{"name":"Nagoya University, Nagoya, Japan"}],"role":[{"role":"author","vocabulary":"crossref"}]}],"member":"320","published-online":{"date-parts":[[2023,10,29]]},"reference":[{"key":"e_1_3_2_1_1_1","volume-title":"Proc. 13th Lang. Resour. Evaluation Conf. (Marseille, Bouches-du-Rh\u00f4ne, France). ELRA","author":"Carlsson Fredrik","year":"2022","unstructured":"Fredrik Carlsson , Philipp Eisen , Faton Rekathati , and Magnus Sahlgren . 2022 . Cross-lingual and multilingual CLIP . In Proc. 13th Lang. Resour. Evaluation Conf. (Marseille, Bouches-du-Rh\u00f4ne, France). ELRA , Paris, France, 6848--6854. Fredrik Carlsson, Philipp Eisen, Faton Rekathati, and Magnus Sahlgren. 2022. Cross-lingual and multilingual CLIP. In Proc. 13th Lang. Resour. Evaluation Conf. (Marseille, Bouches-du-Rh\u00f4ne, France). ELRA, Paris, France, 6848--6854."},{"key":"e_1_3_2_1_2_1","doi-asserted-by":"publisher","DOI":"10.48550\/arXiv.2204.08583"},{"volume-title":"Proc. NeurIPS 2022 Workshop Score-Based Methods","author":"Daras Giannis","key":"e_1_3_2_1_3_1","unstructured":"Giannis Daras and Alexandros G. Dimakis . 2022. Discovering the hidden vocabulary of DALLE-2 . In Proc. NeurIPS 2022 Workshop Score-Based Methods ( New Orleans, LA, USA). bibinfonumpages5 pages. https:\/\/openreview.net\/forum?id=jxeSZaVzpmg Giannis Daras and Alexandros G. Dimakis. 2022. Discovering the hidden vocabulary of DALLE-2. In Proc. NeurIPS 2022 Workshop Score-Based Methods (New Orleans, LA, USA). bibinfonumpages5 pages. https:\/\/openreview.net\/forum?id=jxeSZaVzpmg"},{"key":"e_1_3_2_1_4_1","doi-asserted-by":"publisher","DOI":"10.5220\/0010503701660174"},{"key":"e_1_3_2_1_5_1","doi-asserted-by":"publisher","DOI":"10.1037\/0278--7393.18.6.1211"},{"key":"e_1_3_2_1_6_1","first-page":"9","volume-title":"Adv. Neural Inf. Process. Syst. (Montr\u00e9al, QC, Canada)","author":"Goodfellow Ian J.","unstructured":"Ian J. Goodfellow , Jean Pouget-Abadie , Mehdi Mirza , Bing Xu , David Warde-Farley , Sherjil Ozair , Aaron Courville , and Yoshua Bengio . 2014. Generative Adversarial Nets . In Adv. Neural Inf. Process. Syst. (Montr\u00e9al, QC, Canada) , Vol. 27 . Curran Associates, Inc. , New York, NY, USA , bibinfonumpages 9 pages. Ian J. Goodfellow, Jean Pouget-Abadie, Mehdi Mirza, Bing Xu, David Warde-Farley, Sherjil Ozair, Aaron Courville, and Yoshua Bengio. 2014. Generative Adversarial Nets. In Adv. Neural Inf. Process. Syst. (Montr\u00e9al, QC, Canada), Vol. 27. Curran Associates, Inc., New York, NY, USA, bibinfonumpages9 pages."},{"key":"e_1_3_2_1_7_1","first-page":"12","volume-title":"Adv. Neural Inf. Process. Syst. (Long Beach, CA, USA)","author":"Heusel Martin","unstructured":"Martin Heusel , Hubert Ramsauer , Thomas Unterthiner , Bernhard Nessler , and Sepp Hochreiter . 2017. GANs trained by a two time-scale update rule converge to a local Nash equilibrium . In Adv. Neural Inf. Process. Syst. (Long Beach, CA, USA) , Vol. 30 . Curran Associates, Inc. , New York, NY, USA , bibinfonumpages 12 pages. Martin Heusel, Hubert Ramsauer, Thomas Unterthiner, Bernhard Nessler, and Sepp Hochreiter. 2017. GANs trained by a two time-scale update rule converge to a local Nash equilibrium. In Adv. Neural Inf. Process. Syst. (Long Beach, CA, USA), Vol. 30. Curran Associates, Inc., New York, NY, USA, bibinfonumpages12 pages."},{"key":"e_1_3_2_1_8_1","doi-asserted-by":"publisher","DOI":"10.1017\/CBO9780511751806"},{"key":"e_1_3_2_1_9_1","doi-asserted-by":"publisher","DOI":"10.1145\/3581641.3584078"},{"key":"e_1_3_2_1_10_1","doi-asserted-by":"publisher","DOI":"10.18653\/v1\/W19-4219"},{"key":"e_1_3_2_1_11_1","volume-title":"Learning Multiple Layers of Features from Tiny Images. Master's thesis","author":"Krizhevsky Alex","year":"2009","unstructured":"Alex Krizhevsky . 2009. Learning Multiple Layers of Features from Tiny Images. Master's thesis . University of Toronto , Toronto, ON , Canada. https:\/\/www.cs.toronto.edu\/ kriz\/learning-features- 2009 -TR.pdf Alex Krizhevsky. 2009. Learning Multiple Layers of Features from Tiny Images. Master's thesis. University of Toronto, Toronto, ON, Canada. https:\/\/www.cs.toronto.edu\/ kriz\/learning-features-2009-TR.pdf"},{"key":"e_1_3_2_1_12_1","unstructured":"Wolfgang K\u00f6hler. 1929. Gestalt Psychology. H. Liveright New York NY USA.  Wolfgang K\u00f6hler. 1929. Gestalt Psychology. H. Liveright New York NY USA."},{"key":"e_1_3_2_1_13_1","doi-asserted-by":"publisher","DOI":"10.1007\/978--3--319--10602--1_48"},{"key":"e_1_3_2_1_14_1","doi-asserted-by":"publisher","DOI":"10.48550\/arxiv.2303.03144"},{"key":"e_1_3_2_1_15_1","doi-asserted-by":"publisher","DOI":"10.1037\/h0031564"},{"key":"e_1_3_2_1_16_1","doi-asserted-by":"publisher","DOI":"10.48550\/arXiv.1301.3781"},{"volume-title":"Adv. Neural Inf. Process. Syst. (Lake Tahoe, NV, USA)","author":"Mikolov Tomas","key":"e_1_3_2_1_17_1","unstructured":"Tomas Mikolov , Ilya Sutskever , Kai Chen , Greg Corrado , and Jeffrey Dean . 2013b. Distributed representations of words and phrases and their compositionality . In Adv. Neural Inf. Process. Syst. (Lake Tahoe, NV, USA) , Vol. 26 . Curran Associates, Inc. , New York, NY, USA , 3111--3119. Tomas Mikolov, Ilya Sutskever, Kai Chen, Greg Corrado, and Jeffrey Dean. 2013b. Distributed representations of words and phrases and their compositionality. In Adv. Neural Inf. Process. Syst. (Lake Tahoe, NV, USA), Vol. 26. Curran Associates, Inc., New York, NY, USA, 3111--3119."},{"key":"e_1_3_2_1_18_1","doi-asserted-by":"publisher","DOI":"10.48550\/arXiv.2208.04135"},{"key":"e_1_3_2_1_19_1","volume-title":"Proc. 39th Int. Conf. Mach. Learn., Proc. Mach. Learn. Res.","volume":"162","author":"Nichol Alex","year":"2022","unstructured":"Alex Nichol , Prafulla Dhariwal , Aditya Ramesh , Pranav Shyam , Pamela Mishkin , Bob McGrew , Ilya Sutskever , and Mark Chen . 2022 . GLIDE: Towards photorealistic image generation and editing with text-guided diffusion models . In Proc. 39th Int. Conf. Mach. Learn., Proc. Mach. Learn. Res. ( Baltimore, MD, USA) , Vol. 162 . PMLR, Cambridge, MA, USA, 16784--16804. Alex Nichol, Prafulla Dhariwal, Aditya Ramesh, Pranav Shyam, Pamela Mishkin, Bob McGrew, Ilya Sutskever, and Mark Chen. 2022. GLIDE: Towards photorealistic image generation and editing with text-guided diffusion models. In Proc. 39th Int. Conf. Mach. Learn., Proc. Mach. Learn. Res. (Baltimore, MD, USA), Vol. 162. PMLR, Cambridge, MA, USA, 16784--16804."},{"key":"e_1_3_2_1_20_1","volume-title":"Proc. 38th Int. Conf. Mach. Learn., Proc. Mach. Learn. Res. (Online)","volume":"139","author":"Radford Alec","year":"2021","unstructured":"Alec Radford , Jong Wook Kim , Chris Hallacy , Aditya Ramesh , Gabriel Goh , Sandhini Agarwal , Girish Sastry , Amanda Askell , Pamela Mishkin , Jack Clark , Gretchen Krueger , and Ilya Sutskever . 2021 . Learning transferable visual models from natural language supervision . In Proc. 38th Int. Conf. Mach. Learn., Proc. Mach. Learn. Res. (Online) , Vol. 139 . PMLR, Cambridge, MA, USA, 8748--8763. Alec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh, Gabriel Goh, Sandhini Agarwal, Girish Sastry, Amanda Askell, Pamela Mishkin, Jack Clark, Gretchen Krueger, and Ilya Sutskever. 2021. Learning transferable visual models from natural language supervision. In Proc. 38th Int. Conf. Mach. Learn., Proc. Mach. Learn. Res. (Online), Vol. 139. PMLR, Cambridge, MA, USA, 8748--8763."},{"key":"e_1_3_2_1_21_1","first-page":"3","article-title":"Synaesthesia --A window into perception, thought and language","volume":"8","author":"Ramachandran Vilayanur S.","year":"2001","unstructured":"Vilayanur S. Ramachandran and Edward M. Hubbard . 2001 . Synaesthesia --A window into perception, thought and language . J. Conscious. Stud. , Vol. 8 , 12 (2001), 3 -- 34 . Vilayanur S. Ramachandran and Edward M. Hubbard. 2001. Synaesthesia --A window into perception, thought and language. J. Conscious. Stud. , Vol. 8, 12 (2001), 3--34.","journal-title":"J. Conscious. Stud."},{"key":"e_1_3_2_1_22_1","doi-asserted-by":"publisher","DOI":"10.48550\/arXiv.2204.06125"},{"key":"e_1_3_2_1_23_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR52688.2022.01042"},{"key":"e_1_3_2_1_24_1","doi-asserted-by":"publisher","DOI":"10.18653\/v1"},{"key":"e_1_3_2_1_25_1","doi-asserted-by":"publisher","DOI":"10.48550\/arxiv.2205.11487"},{"key":"e_1_3_2_1_26_1","volume-title":"Proc. 36th Conf. Neural Inf. Process. Syst. Datasets Benchmarks Track","author":"Schuhmann Christoph","year":"2022","unstructured":"Christoph Schuhmann , Romain Beaumont , Richard Vencu , Cade W. Gordon , Ross Wightman , Mehdi Cherti , Theo Coombes , Aarush Katta , Clayton Mullis , Mitchell Wortsman , Patrick Schramowski , Srivatsa R. Kundurthy , Katherine Crowson , Ludwig Schmidt , Robert Kaczmarczyk , and Jenia Jitsev . 2022 . LAION-5B: An open large-scale dataset for training next generation image-text models . In Proc. 36th Conf. Neural Inf. Process. Syst. Datasets Benchmarks Track ( New Orleans, LA, USA). Curran Associates, Inc., New York, NY, USA, bibinfonumpages17 pages. Christoph Schuhmann, Romain Beaumont, Richard Vencu, Cade W. Gordon, Ross Wightman, Mehdi Cherti, Theo Coombes, Aarush Katta, Clayton Mullis, Mitchell Wortsman, Patrick Schramowski, Srivatsa R. Kundurthy, Katherine Crowson, Ludwig Schmidt, Robert Kaczmarczyk, and Jenia Jitsev. 2022. LAION-5B: An open large-scale dataset for training next generation image-text models. In Proc. 36th Conf. Neural Inf. Process. Syst. Datasets Benchmarks Track (New Orleans, LA, USA). Curran Associates, Inc., New York, NY, USA, bibinfonumpages17 pages."},{"key":"e_1_3_2_1_27_1","doi-asserted-by":"publisher","DOI":"10.18653\/v1\/P16-1162"},{"key":"e_1_3_2_1_28_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR52688.2022.00939"},{"key":"e_1_3_2_1_29_1","volume-title":"Proc. 32nd Int. Conf. Mach. Learn., Proc. Mach. Learn. Res. (Lille","volume":"37","author":"Sohl-Dickstein Jascha","year":"2015","unstructured":"Jascha Sohl-Dickstein , Eric Weiss , Niru Maheswaranathan , and Surya Ganguli . 2015 . Deep unsupervised learning using nonequilibrium thermodynamics . In Proc. 32nd Int. Conf. Mach. Learn., Proc. Mach. Learn. Res. (Lille , Nord, France) , Vol. 37 . PMLR, Cambridge, MA, USA, 2256--2265. Jascha Sohl-Dickstein, Eric Weiss, Niru Maheswaranathan, and Surya Ganguli. 2015. Deep unsupervised learning using nonequilibrium thermodynamics. In Proc. 32nd Int. Conf. Mach. Learn., Proc. Mach. Learn. Res. (Lille, Nord, France), Vol. 37. PMLR, Cambridge, MA, USA, 2256--2265."},{"key":"e_1_3_2_1_30_1","doi-asserted-by":"publisher","DOI":"10.5281\/zenodo.1443582"},{"key":"e_1_3_2_1_31_1","volume-title":"https:\/\/github.com\/CompVis\/stable-diffusion\/ (Accessed","author":"Computer Vision and Learning Research Group at Ludwig Maximilian University of Munich. 2022. Stable Diffusion.","year":"2023","unstructured":"Computer Vision and Learning Research Group at Ludwig Maximilian University of Munich. 2022. Stable Diffusion. https:\/\/github.com\/CompVis\/stable-diffusion\/ (Accessed July 11, 2023 ). Computer Vision and Learning Research Group at Ludwig Maximilian University of Munich. 2022. Stable Diffusion. https:\/\/github.com\/CompVis\/stable-diffusion\/ (Accessed July 11, 2023)."},{"key":"e_1_3_2_1_32_1","doi-asserted-by":"publisher","DOI":"10.1109\/ICASSP43922.2022.9747669"}],"event":{"name":"MM '23: The 31st ACM International Conference on Multimedia","sponsor":["SIGMM ACM Special Interest Group on Multimedia"],"location":"Ottawa ON Canada","acronym":"MM '23"},"container-title":["Proceedings of the 1st International Workshop on Multimedia Content Generation and Evaluation: New Methods and Practice"],"original-title":[],"link":[{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/3607541.3616818","content-type":"unspecified","content-version":"vor","intended-application":"text-mining"},{"URL":"https:\/\/dl.acm.org\/doi\/pdf\/10.1145\/3607541.3616818","content-type":"unspecified","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2025,6,17]],"date-time":"2025-06-17T16:36:28Z","timestamp":1750178188000},"score":1,"resource":{"primary":{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/3607541.3616818"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2023,10,29]]},"references-count":32,"alternative-id":["10.1145\/3607541.3616818","10.1145\/3607541"],"URL":"https:\/\/doi.org\/10.1145\/3607541.3616818","relation":{},"subject":[],"published":{"date-parts":[[2023,10,29]]},"assertion":[{"value":"2023-10-29","order":2,"name":"published","label":"Published","group":{"name":"publication_history","label":"Publication History"}}]}}