{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2026,6,9]],"date-time":"2026-06-09T15:00:28Z","timestamp":1781017228120,"version":"3.54.1"},"reference-count":45,"publisher":"Springer Science and Business Media LLC","issue":"20","license":[{"start":{"date-parts":[[2022,4,4]],"date-time":"2022-04-04T00:00:00Z","timestamp":1649030400000},"content-version":"tdm","delay-in-days":0,"URL":"https:\/\/creativecommons.org\/licenses\/by\/4.0"},{"start":{"date-parts":[[2022,4,4]],"date-time":"2022-04-04T00:00:00Z","timestamp":1649030400000},"content-version":"vor","delay-in-days":0,"URL":"https:\/\/creativecommons.org\/licenses\/by\/4.0"}],"content-domain":{"domain":["link.springer.com"],"crossmark-restriction":false},"short-container-title":["Multimed Tools Appl"],"published-print":{"date-parts":[[2022,8]]},"abstract":"<jats:title>Abstract<\/jats:title><jats:p>Cosplay has grown from its origins at fan conventions into a billion-dollar global dress phenomenon. To facilitate the imagination and reinterpretation of animated images as real garments, this paper presents an automatic costume-image generation method based on image-to-image translation. Cosplay items can be significantly diverse in their styles and shapes, and conventional methods cannot be directly applied to the wide variety of clothing images that are the focus of this study. To solve this problem, our method starts by collecting and preprocessing web images to prepare a cleaned, paired dataset of the anime and real domains. Then, we present a novel architecture for generative adversarial networks (GANs) to facilitate high-quality cosplay image generation. Our GAN consists of several effective techniques to bridge the two domains and improve both the global and local consistency of generated images. Experiments demonstrated that, with quantitative evaluation metrics, the proposed GAN performs better and produces more realistic images than conventional methods. Our codes and pretrained model are available on the web.<\/jats:p>","DOI":"10.1007\/s11042-022-12576-x","type":"journal-article","created":{"date-parts":[[2022,4,4]],"date-time":"2022-04-04T18:03:35Z","timestamp":1649095415000},"page":"29505-29523","update-policy":"https:\/\/doi.org\/10.1007\/springer_crossmark_policy","source":"Crossref","is-referenced-by-count":8,"title":["Anime-to-real clothing: Cosplay costume generation via image-to-image translation"],"prefix":"10.1007","volume":"81","author":[{"given":"Koya","family":"Tango","sequence":"first","affiliation":[],"role":[{"vocabulary":"crossref","role":"author"}]},{"ORCID":"https:\/\/orcid.org\/0000-0003-4899-2427","authenticated-orcid":false,"given":"Marie","family":"Katsurai","sequence":"additional","affiliation":[],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Hayato","family":"Maki","sequence":"additional","affiliation":[],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Ryosuke","family":"Goto","sequence":"additional","affiliation":[],"role":[{"vocabulary":"crossref","role":"author"}]}],"member":"297","published-online":{"date-parts":[[2022,4,4]]},"reference":[{"key":"12576_CR1","unstructured":"Azathoth (2018) MyAnimeList Dataset: contains 300k users, 14k anime metadata, and 80mil. ratings from MyAnimeList.net. https:\/\/www.kaggle.com\/azathoth42\/myanimelist. Accessed 19 Aug 2020"},{"key":"12576_CR2","doi-asserted-by":"crossref","unstructured":"Chen Y, Wang Z, Peng Y, Zhang Z, Yu G, Sun J (2018) Cascaded Pyramid Network for Multi-Person Pose Estimation. In: Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition, pp 7103\u20137112","DOI":"10.1109\/CVPR.2018.00742"},{"key":"12576_CR3","unstructured":"Cheng W-H, Song S, Chen C-Y, Hidayati SC, Liu J (2020) Fashion Meets Computer Vision: A Survey. arXiv:2003.13988"},{"key":"12576_CR4","doi-asserted-by":"crossref","unstructured":"Ci Y, Ma X, Wang Z, Li H, Luo Z (2018) User-guided deep anime line art colorization with conditional adversarial networks. In: Proceedings of the 26th ACM international conference on multimedia, pp 1536\u20131544","DOI":"10.1145\/3240508.3240661"},{"key":"12576_CR5","unstructured":"CooperUnion (2016) Anime Recommendations Database: Recommendation data from 76,000 users at myanimelist.net. https:\/\/www.kaggle.com\/CooperUnion\/anime-recommendations-database. Accessed 19 Aug 2020"},{"key":"12576_CR6","doi-asserted-by":"crossref","unstructured":"Cordts M, Omran M, Ramos S, Rehfeld T, Enzweiler M, Benenson R, Franke U, Roth S, Schiele B (2016) The Cityscapes Dataset for Semantic Urban Scene Understanding. In: Proceedings of the IEEE Conferene on Computer Vision and Pattern Recognition, pp 3213\u20133223","DOI":"10.1109\/CVPR.2016.350"},{"key":"12576_CR7","unstructured":"Goodfellow I, Pouget-Abadie J, Mirza M, Xu B, Warde-Farley D, Ozair S, Courville A, Bengio Y (2014) Generative Adversarial Nets. In: Advances in Neural Information Processing Systems, pp 2672\u20132680"},{"key":"12576_CR8","doi-asserted-by":"crossref","unstructured":"Hamada K, Tachibana K, Li T, Honda H, Uchida Y (2018) Full-Body High-Resolution Anime Generation with Progressive Structure-Conditional Generative Adversarial Networks. In: European Conference on Computer Vision. Springer, pp 67\u201374","DOI":"10.1007\/978-3-030-11015-4_8"},{"key":"12576_CR9","doi-asserted-by":"crossref","unstructured":"Han X, Wu Z, Wu Z, Yu R, Davis L S (2018) Viton: An Image-based Virtual Try-on Network. In: Proceedings of the IEEE Conferene on Computer Vision and Pattern Recognition, pp 7543\u20137552","DOI":"10.1109\/CVPR.2018.00787"},{"key":"12576_CR10","unstructured":"Heusel M, Ramsauer H, Unterthiner T, Nessler B, Hochreiter S (2017) GANs Trained by a Two Time-Scale Update Rule Converge to a Local Nash Equilibrium. In: Advances in Neural Information Processing Systems, pp 6626\u20136637"},{"key":"12576_CR11","doi-asserted-by":"crossref","unstructured":"Isola P, Zhu J-Y, Zhou T, Efros A A (2017) Image-to-Image Translation with Conditional Adversarial Networks. In: Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition, pp 1125\u20131134","DOI":"10.1109\/CVPR.2017.632"},{"key":"12576_CR12","unstructured":"Jin Y, Zhang J, Li M, Tian Y, Zhu H, Fang Z (2017) Towards the Automatic Anime Characters Creation with Generative Adversarial Networks. arXiv:1708.05509"},{"key":"12576_CR13","unstructured":"Karras T, Aila T, Laine S, Lehtinen J (2017) Progressive Growing of GANs for Improved Quality, Stability, and Variation. arXiv:1710.10196"},{"key":"12576_CR14","unstructured":"Kingma DP, Ba J (2014) Adam: A Method for Stochastic Optimization. arXiv:1412.6980"},{"key":"12576_CR15","unstructured":"Krizhevsky A, Sutskever I, Hinton G E (2012) ImageNet Classification with Deep Convolutional Neural Networks. In: Advances in Neural Information Processing Systems, pp 1097\u20131105"},{"key":"12576_CR16","doi-asserted-by":"crossref","unstructured":"Kwon Y, Kim S, Yoo D, Yoon S-E (2019) Coarse-to-Fine Clothing Image Generation with Progressively Constructed Conditional GAN. In: 14th International Conference on Computer Vision Theory and Applications, SCITEPRESS-Science and Technology Publications, pp 83\u201390","DOI":"10.5220\/0007306900830090"},{"key":"12576_CR17","unstructured":"Li V (2018) FashionAI KeyPoint Detection Challenge Keras. https:\/\/github.com\/yuanyuanli85\/FashionAI_KeyPoint_Detection_Challenge_Keras. Accessed 19 Aug 2020"},{"key":"12576_CR18","doi-asserted-by":"crossref","unstructured":"Liu W, Anguelov D, Erhan D, Szegedy C, Reed S, Fu C-Y, Berg AC (2016) SSD: Single Shot Multibox Detector. In: European Conference on Computer Vision. Springer, pp 21\u201337","DOI":"10.1007\/978-3-319-46448-0_2"},{"key":"12576_CR19","doi-asserted-by":"crossref","unstructured":"Liu Z, Luo P, Qiu S, X, Tang X (2016) DeepFashion: Powering robust clothes recognition and retrieval with rich annotations. In: Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition (CVPR), pp 1096\u20131104","DOI":"10.1109\/CVPR.2016.124"},{"key":"12576_CR20","unstructured":"lltcggie (2018) Waifu2x-Caffe. https:\/\/github.com\/lltcggie\/waifu2x-caffe. (Last accessed: 19\/08\/2020)"},{"key":"12576_CR21","doi-asserted-by":"crossref","unstructured":"Long J, Shelhamer E, Darrell T (2015) Fully Convolutional Networks for Semantic Segmentation. In: Proceedings of the IEEE Conferene on Computer Vision and Pattern Recognition, pp 3431\u20133440","DOI":"10.1109\/CVPR.2015.7298965"},{"key":"12576_CR22","doi-asserted-by":"crossref","unstructured":"Mao X, Li Q, Xie H, Lau RY, Wang Z, Paul Smolley S (2017) Least Squares Generative Adversarial Networks. In: Proceedings of the IEEE International Conference on Computer Vision, pp 2794\u20132802","DOI":"10.1109\/ICCV.2017.304"},{"key":"12576_CR23","unstructured":"Mirza M, Osindero S (2014) Conditional Generative Adversarial Nets. arXiv:1411.1784"},{"key":"12576_CR24","unstructured":"Miyato T, Kataoka T, Koyama M, Yoshida Y (2018) Spectral Normalization for Generative Adversarial Networks. arXiv:1802.05957"},{"key":"12576_CR25","doi-asserted-by":"crossref","unstructured":"Park T, Liu M-Y, T-C, Zhu J-Y (2019) Semantic Image Synthesis with Spatially-Adaptive Normalization. In: Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition, pp 2337\u20132346","DOI":"10.1109\/CVPR.2019.00244"},{"key":"12576_CR26","unstructured":"ripobi-tan (2016) DupFileEliminator. https:\/\/www.vector.co.jp\/soft\/winnt\/util\/se492140.html. Accessed 19 Aug 2020"},{"key":"12576_CR27","doi-asserted-by":"crossref","unstructured":"Ronneberger O, Fischer P, Brox T (2015) U-Net: Convolutional Networks for Biomedical Image Segmentation. In: International Conference on Medical Image Computing and Computer-Assisted Intervention. Springer, pp 234\u2013241","DOI":"10.1007\/978-3-319-24574-4_28"},{"key":"12576_CR28","doi-asserted-by":"crossref","unstructured":"Royer A, Bousmalis K, Gouws S, Bertsch F, Mosseri I, Cole F, Murphy K (2020) XGAN: Unsupervised Image-to-Image Translation for Many-to-Many Mappings. In: Domain Adaptation for Visual Understanding. Springer, pp 33\u201349","DOI":"10.1007\/978-3-030-30671-7_3"},{"key":"12576_CR29","unstructured":"Salimans T, Goodfellow I, Zaremba W, Cheung V, Radford A, Chen X (2016) Improved Techniques for Training GANs. In: Advances in Neural Information Processing Systems, pp 2234\u20132242"},{"key":"12576_CR30","doi-asserted-by":"crossref","unstructured":"Shocher A, Bagon S, Isola P, Irani M (2018) InGAN: Capturing and Remapping the ``DNA\u201d of a Natural Image. arXiv:1812.00231","DOI":"10.1109\/ICCV.2019.00459"},{"issue":"1","key":"12576_CR31","doi-asserted-by":"publisher","first-page":"60","DOI":"10.1186\/s40537-019-0197-0","volume":"6","author":"C Shorten","year":"2019","unstructured":"Shorten C, Khoshgoftaar TM (2019) A Survey on Image Data Augmentation for Deep Learning. J Big Data 6(1):60","journal-title":"J Big Data"},{"key":"12576_CR32","unstructured":"Simonyan K, Zisserman A (2014) Very Deep Convolutional Networks for Large-Scale Image Recognition. arXiv:1409.1556"},{"key":"12576_CR33","doi-asserted-by":"crossref","unstructured":"Szegedy C, Vanhoucke V, Ioffe S, Shlens J, Wojna Z (2016) Rethinking the Inception Architecture for Computer Vision. In: Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition, pp 2818\u20132826","DOI":"10.1109\/CVPR.2016.308"},{"key":"12576_CR34","doi-asserted-by":"crossref","unstructured":"Tang H, Xu D, Sebe N, Yan Y (2019) Attention-Guided Generative Adversarial Networks for Unsupervised Image-to-Image Translation. arXiv:1903.12296","DOI":"10.1109\/IJCNN.2019.8851881"},{"key":"12576_CR35","doi-asserted-by":"crossref","unstructured":"Ulyanov D, Vedaldi A, Lempitsky V (2017) Improved Texture Networks: Maximizing Quality and Diversity in Feed-forward Stylization and Texture Synthesis. In: Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition, pp 6924\u20136932","DOI":"10.1109\/CVPR.2017.437"},{"issue":"1","key":"12576_CR36","doi-asserted-by":"publisher","first-page":"24","DOI":"10.1007\/s11263-010-0372-4","volume":"91","author":"S Vijayanarasimhan","year":"2011","unstructured":"Vijayanarasimhan S, Grauman K (2011) Cost-Sensitive Active Visual Category Learning. Int J Comput Vis 91(1):24\u201344","journal-title":"Int J Comput Vis"},{"key":"12576_CR37","doi-asserted-by":"crossref","unstructured":"Wang T-C, Liu M-Y, Zhu J-Y, Tao A, Kautz J, Catanzaro B (2018) High-Resolution Image Synthesis and Semantic Manipulation with Conditional GANs. In: Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition, pp 8798\u20138807","DOI":"10.1109\/CVPR.2018.00917"},{"key":"12576_CR38","unstructured":"WCS Inc. (2019) What\u2019s WCS?. https:\/\/en.worldcosplaysummit.jp\/championship2019-about. 24 May 2020"},{"key":"12576_CR39","doi-asserted-by":"crossref","unstructured":"Wu W, Cao K, Li C, Qian C, Loy CC (2019) Transgaga: Geometry-Aware Unsupervised Image-to-Image Translation. In: Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition, pp 8012\u20138021","DOI":"10.1109\/CVPR.2019.00820"},{"key":"12576_CR40","doi-asserted-by":"crossref","unstructured":"Wu Z, Lin G, Tao Q, Cai J (2019) M2E-Try On Net: Fashion from Model to Everyone. In: Proceedings of the 27th ACM International Conference on Multimedia, pp 293\u2013301","DOI":"10.1145\/3343031.3351083"},{"key":"12576_CR41","doi-asserted-by":"crossref","unstructured":"Yoo D, Kim N, Park S, Paek AS, Kweon IS (2016) Pixel-Level Domain Transfer. In: European Conference on Computer Vision. Springer, pp 517\u2013532","DOI":"10.1007\/978-3-319-46484-8_31"},{"key":"12576_CR42","doi-asserted-by":"crossref","unstructured":"Zhang R, Isola P, Efros AA, Shechtman E, Wang O (2018) The Unreasonable Effectiveness of Deep Features as a Perceptual Metric. In: Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition, pp 586\u2013595","DOI":"10.1109\/CVPR.2018.00068"},{"key":"12576_CR43","doi-asserted-by":"crossref","unstructured":"Zhu J-Y, Park T, Isola P, Efros AA (2017) Unpaired Image-to-Image Translation using Cycle-Consistent Adversarial Networks. In: Proceedings of the IEEE International Conference on Computer Vision, pp 2223\u20132232","DOI":"10.1109\/ICCV.2017.244"},{"key":"12576_CR44","doi-asserted-by":"crossref","unstructured":"Zhu S, Urtasun R, Fidler S, Lin D, Change Loy C (2017) Be Your Own Prada: Fashion Synthesis with Structural Coherence. In: Proceedings of the IEEE International Conference on Computer Vision, pp 1680\u20131688","DOI":"10.1109\/ICCV.2017.186"},{"key":"12576_CR45","doi-asserted-by":"crossref","unstructured":"Zou X, Kong X, Wong W, Wang C, Liu Y, Cao Y (2019) FashionAI: A Hierarchical Dataset for Fashion Understanding. In: Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition Workshops, pp 296\u2013304","DOI":"10.1109\/CVPRW.2019.00039"}],"container-title":["Multimedia Tools and Applications"],"original-title":[],"language":"en","link":[{"URL":"https:\/\/link.springer.com\/content\/pdf\/10.1007\/s11042-022-12576-x.pdf","content-type":"application\/pdf","content-version":"vor","intended-application":"text-mining"},{"URL":"https:\/\/link.springer.com\/article\/10.1007\/s11042-022-12576-x\/fulltext.html","content-type":"text\/html","content-version":"vor","intended-application":"text-mining"},{"URL":"https:\/\/link.springer.com\/content\/pdf\/10.1007\/s11042-022-12576-x.pdf","content-type":"application\/pdf","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2022,7,23]],"date-time":"2022-07-23T07:21:15Z","timestamp":1658560875000},"score":1,"resource":{"primary":{"URL":"https:\/\/link.springer.com\/10.1007\/s11042-022-12576-x"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2022,4,4]]},"references-count":45,"journal-issue":{"issue":"20","published-print":{"date-parts":[[2022,8]]}},"alternative-id":["12576"],"URL":"https:\/\/doi.org\/10.1007\/s11042-022-12576-x","relation":{},"ISSN":["1380-7501","1573-7721"],"issn-type":[{"value":"1380-7501","type":"print"},{"value":"1573-7721","type":"electronic"}],"subject":[],"published":{"date-parts":[[2022,4,4]]},"assertion":[{"value":"26 August 2020","order":1,"name":"received","label":"Received","group":{"name":"ArticleHistory","label":"Article History"}},{"value":"28 February 2021","order":2,"name":"revised","label":"Revised","group":{"name":"ArticleHistory","label":"Article History"}},{"value":"31 January 2022","order":3,"name":"accepted","label":"Accepted","group":{"name":"ArticleHistory","label":"Article History"}},{"value":"4 April 2022","order":4,"name":"first_online","label":"First Online","group":{"name":"ArticleHistory","label":"Article History"}}]}}