{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2026,6,4]],"date-time":"2026-06-04T06:44:37Z","timestamp":1780555477241,"version":"3.54.1"},"reference-count":34,"publisher":"MDPI AG","issue":"7","license":[{"start":{"date-parts":[[2022,3,30]],"date-time":"2022-03-30T00:00:00Z","timestamp":1648598400000},"content-version":"vor","delay-in-days":0,"URL":"https:\/\/creativecommons.org\/licenses\/by\/4.0\/"}],"content-domain":{"domain":[],"crossmark-restriction":false},"short-container-title":["Sensors"],"abstract":"<jats:p>Consumer-to-shop clothes retrieval refers to the problem of matching photos taken by customers with their counterparts in the shop. Due to some problems, such as a large number of clothing categories, different appearances of clothing items due to different camera angles and shooting conditions, different background environments, and different body postures, the retrieval accuracy of traditional consumer-to-shop models is always low. With advances in convolutional neural networks (CNNs), the accuracy of garment retrieval has been significantly improved. Most approaches addressing this problem use single CNNs in conjunction with a softmax loss function to extract discriminative features. In the fashion domain, negative pairs can have small or large visual differences that make it difficult to minimize intraclass variance and maximize interclass variance with softmax. Margin-based softmax losses such as Additive Margin-Softmax (aka CosFace) improve the discriminative power of the original softmax loss, but since they consider the same margin for the positive and negative pairs, they are not suitable for cross-domain fashion search. In this work, we introduce the cross-domain discriminative margin loss (DML) to deal with the large variability of negative pairs in fashion. DML learns two different margins for positive and negative pairs such that the negative margin is larger than the positive margin, which provides stronger intraclass reduction for negative pairs. The experiments conducted on publicly available fashion datasets DARN and two benchmarks of the DeepFashion dataset\u2014(1) Consumer-to-Shop Clothes Retrieval and (2) InShop Clothes Retrieval\u2014confirm that the proposed loss function not only outperforms the existing loss functions but also achieves the best performance.<\/jats:p>","DOI":"10.3390\/s22072660","type":"journal-article","created":{"date-parts":[[2022,3,30]],"date-time":"2022-03-30T21:28:39Z","timestamp":1648675719000},"page":"2660","update-policy":"https:\/\/doi.org\/10.3390\/mdpi_crossmark_policy","source":"Crossref","is-referenced-by-count":16,"title":["Deep Learning with Discriminative Margin Loss for Cross-Domain Consumer-to-Shop Clothes Retrieval"],"prefix":"10.3390","volume":"22","author":[{"given":"Pendar","family":"Alirezazadeh","sequence":"first","affiliation":[{"name":"Department of Informatics, University of the Basque Country, 20008 Donostia-San Sebastian, Spain"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"ORCID":"https:\/\/orcid.org\/0000-0001-6581-9680","authenticated-orcid":false,"given":"Fadi","family":"Dornaika","sequence":"additional","affiliation":[{"name":"Department of Informatics, University of the Basque Country, 20008 Donostia-San Sebastian, Spain"},{"name":"School of Computer and Information Engineering, Henan University, Kaifeng 475000, China"},{"name":"Ikerbasque, Basque Foundation for Science, Plaza Euskadi, 5, 48009 Bilbao, Spain"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Abdelmalik","family":"Moujahid","sequence":"additional","affiliation":[{"name":"Department of Mathematics, University of the Basque Country, 48080 Bilbao, Spain"}],"role":[{"vocabulary":"crossref","role":"author"}]}],"member":"1968","published-online":{"date-parts":[[2022,3,30]]},"reference":[{"key":"ref_1","doi-asserted-by":"crossref","unstructured":"Hadi Kiapour, M., Han, X., Lazebnik, S., Berg, A., and Berg, T. (2015, January 7\u201313). Where to buy it: Matching street clothing photos in online shops. Proceedings of the IEEE International Conference on Computer Vision, Santiago, Chile.","DOI":"10.1109\/ICCV.2015.382"},{"key":"ref_2","doi-asserted-by":"crossref","unstructured":"Li, Z., Li, Y., Gao, Y., and Liu, Y. (2016). Fast cross-scenario clothing retrieval based on indexing deep features. Pacific Rim Conference on Multimedia, Springer.","DOI":"10.1007\/978-3-319-48890-5_11"},{"key":"ref_3","doi-asserted-by":"crossref","unstructured":"Liu, S., Song, Z., Liu, G., Xu, C., Lu, H., and Yan, S. (2012, January 16\u201321). Street-to-shop: Cross-scenario clothing retrieval via parts alignment and auxiliary set. Proceedings of the 2012 IEEE Conference on Computer Vision and Pattern Recognition, Providence, RI, USA.","DOI":"10.1145\/2393347.2396471"},{"key":"ref_4","doi-asserted-by":"crossref","unstructured":"Wang, X., Sun, Z., Zhang, W., Zhou, Y., and Jiang, Y. (2016, January 6\u20139). Matching user photos to online products with robust deep features. Proceedings of the 2016 ACM on International Conference on Multimedia Retrieval, New York, NY, USA.","DOI":"10.1145\/2911996.2912002"},{"key":"ref_5","doi-asserted-by":"crossref","unstructured":"Kalantidis, Y., Kennedy, L., and Li, L. (2013, January 16\u201320). Getting the look: Clothing recognition and segmentation for automatic product suggestions in everyday photos. Proceedings of the 3rd ACM Conference on International Conference on Multimedia Retrieval, Dallas, TX, USA.","DOI":"10.1145\/2461466.2461485"},{"key":"ref_6","doi-asserted-by":"crossref","unstructured":"Ji, X., Wang, W., Zhang, M., and Yang, Y. (2017, January 23\u201327). Cross-domain image retrieval with attention modeling. Proceedings of the 25th ACM International Conference on Multimedia, Mountain View, CA, USA.","DOI":"10.1145\/3123266.3123429"},{"key":"ref_7","doi-asserted-by":"crossref","unstructured":"Cheng, Z., Wu, X., Liu, Y., and Hua, X. (2017, January 21\u201326). Video2shop: Exact matching clothes in videos to online shopping images. Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition, Honolulu, HI, USA.","DOI":"10.1109\/CVPR.2017.444"},{"key":"ref_8","doi-asserted-by":"crossref","unstructured":"Wang, Z., Gu, Y., Zhang, Y., Zhou, J., and Gu, X. (2017, January 10\u201313). Clothing retrieval with visual attention model. Proceedings of the 2017 IEEE Visual Communications and Image Processing (VCIP), St. Petersburg, FL, USA.","DOI":"10.1109\/VCIP.2017.8305144"},{"key":"ref_9","unstructured":"Lasserre, J., Bracher, C., and Vollgraf, R. Street2Fashion2Shop: Enabling Visual Search in Fashion e-Commerce Using Studio Images. Proceedings of the International Conference on Pattern Recognition Applications and Methods."},{"key":"ref_10","doi-asserted-by":"crossref","unstructured":"Gajic, B., and Baldrich, R. (2018, January 18\u201322). Cross-domain fashion image retrieval. Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition Workshops, Salt Lake City, UT, USA.","DOI":"10.1109\/CVPRW.2018.00243"},{"key":"ref_11","doi-asserted-by":"crossref","unstructured":"Kuang, Z., Gao, Y., Li, G., Luo, P., Chen, Y., Lin, L., and Zhang, W. (2019). Fashion Retrieval via Graph Reasoning Networks on a Similarity Pyramid. arXiv.","DOI":"10.1109\/ICCV.2019.00316"},{"key":"ref_12","doi-asserted-by":"crossref","unstructured":"Park, S., Shin, M., Ham, S., Choe, S., and Kang, Y. (2019, January 16\u201317). Study on Fashion Image Retrieval Methods for Efficient Fashion Visual Search. Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition Workshops, Long Beach, CA, USA.","DOI":"10.1109\/CVPRW.2019.00042"},{"key":"ref_13","doi-asserted-by":"crossref","unstructured":"Kucer, M., and Murray, N. (2019, January 16\u201317). A Detect-Then-Retrieve Model for Multi-Domain Fashion Item Retrieval. Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition Workshops, Long Beach, CA, USA.","DOI":"10.1109\/CVPRW.2019.00047"},{"key":"ref_14","doi-asserted-by":"crossref","unstructured":"Chopra, A., Sinha, A., Gupta, H., Sarkar, M., Ayush, K., and Krishnamurthy, B. (2019, January 16\u201317). Powering Robust Fashion Retrieval With Information Rich Feature Embeddings. Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition Workshops, Long Beach, CA, USA.","DOI":"10.1109\/CVPRW.2019.00045"},{"key":"ref_15","doi-asserted-by":"crossref","first-page":"142669","DOI":"10.1109\/ACCESS.2020.3013631","article-title":"ClothingNet: Cross-Domain Clothing Retrieval With Feature Fusion and Quadruplet Loss","volume":"8","author":"Miao","year":"2020","journal-title":"IEEE Access"},{"key":"ref_16","doi-asserted-by":"crossref","unstructured":"Liu, Z., Luo, P., Qiu, S., Wang, X., and Tang, X. (2016, January 27\u201330). Deepfashion: Powering robust clothes recognition and retrieval with rich annotations. Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition, Las Vegas, NV, USA.","DOI":"10.1109\/CVPR.2016.124"},{"key":"ref_17","doi-asserted-by":"crossref","unstructured":"Huang, J., Feris, R., Chen, Q., and Yan, S. (2015, January 7\u201313). Cross-domain image retrieval with a dual attribute-aware ranking network. Proceedings of the IEEE International Conference on Computer Vision, Santiago, Chile.","DOI":"10.1109\/ICCV.2015.127"},{"key":"ref_18","doi-asserted-by":"crossref","unstructured":"Wang, H., Wang, Y., Zhou, Z., Ji, X., Gong, D., Zhou, J., Li, Z., and Liu, W. (2018, January 18\u201322). Cosface: Large margin cosine loss for deep face recognition. Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition, Salt Lake City, UT, USA.","DOI":"10.1109\/CVPR.2018.00552"},{"key":"ref_19","unstructured":"Chopra, S., Hadsell, R., and LeCun, Y. (2005, January 20\u201325). Learning a similarity metric discriminatively, with application to face verification. Proceedings of the 2005 IEEE Computer Society Conference on Computer Vision and Pattern Recognition (CVPR\u201905), San Diego, CA, USA."},{"key":"ref_20","unstructured":"Hadsell, R., Chopra, S., and LeCun, Y. (2006, January 17\u201322). Dimensionality reduction by learning an invariant mapping. Proceedings of the 2006 IEEE Computer Society Conference on Computer Vision and Pattern Recognition (CVPR\u201906), New York, NY, USA."},{"key":"ref_21","doi-asserted-by":"crossref","first-page":"701","DOI":"10.1007\/s11263-018-1135-x","article-title":"Learning Discriminative Aggregation Network for Video-Based Face Recognition and Person Re-identification","volume":"127","author":"Rao","year":"2019","journal-title":"Int. J. Comput. Vis."},{"key":"ref_22","doi-asserted-by":"crossref","unstructured":"Wang, J., Song, Y., Leung, T., Rosenberg, C., Wang, J., Philbin, J., Chen, B., and Wu, Y. (2014, January 23\u201328). Learning fine-grained image similarity with deep ranking. Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition, Columbus, OH, USA.","DOI":"10.1109\/CVPR.2014.180"},{"key":"ref_23","doi-asserted-by":"crossref","unstructured":"Liu, W., Wen, Y., Yu, Z., Li, M., Raj, B., and Song, L. (2017, January 21\u201326). Sphereface: Deep hypersphere embedding for face recognition. Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition, Honolulu, HI, USA.","DOI":"10.1109\/CVPR.2017.713"},{"key":"ref_24","doi-asserted-by":"crossref","unstructured":"Wen, Y., Zhang, K., Li, Z., and Qiao, Y. (2016, January 11\u201314). A discriminative feature learning approach for deep face recognition. Proceedings of the European Conference on Computer Vision, Amsterdam, The Netherlands.","DOI":"10.1007\/978-3-319-46478-7_31"},{"key":"ref_25","first-page":"7","article-title":"Large-margin softmax loss for convolutional neural networks","volume":"2","author":"Liu","year":"2016","journal-title":"ICML"},{"key":"ref_26","doi-asserted-by":"crossref","first-page":"926","DOI":"10.1109\/LSP.2018.2822810","article-title":"Additive margin softmax for face verification","volume":"25","author":"Wang","year":"2018","journal-title":"IEEE Signal Process. Lett."},{"key":"ref_27","doi-asserted-by":"crossref","unstructured":"Deng, J., Guo, J., Xue, N., and Zafeiriou, S. (2019, January 16\u201320). Arcface: Additive angular margin loss for deep face recognition. Proceedings of the IEEE\/CVF Conference on Computer Vision and Pattern Recognition, Long Beach, CA, USA.","DOI":"10.1109\/CVPR.2019.00482"},{"key":"ref_28","doi-asserted-by":"crossref","unstructured":"Hoffer, E., and Ailon, N. (2015). Deep metric learning using triplet network. International Workshop on Similarity-Based Pattern Recognition, Springer.","DOI":"10.1007\/978-3-319-24261-3_7"},{"key":"ref_29","doi-asserted-by":"crossref","first-page":"27","DOI":"10.1016\/j.patrec.2021.01.010","article-title":"Deep learning for real-time semantic segmentation: Application in ultrasound imaging","volume":"144","author":"Ouahabi","year":"2021","journal-title":"Pattern Recognit. Lett."},{"key":"ref_30","doi-asserted-by":"crossref","unstructured":"Xuan, H., Souvenir, R., and Pless, R. (2018, January 8\u201314). Deep randomized ensembles for metric learning. Proceedings of the European Conference on Computer Vision (ECCV), Munich, Germany.","DOI":"10.1007\/978-3-030-01270-0_44"},{"key":"ref_31","doi-asserted-by":"crossref","unstructured":"Shen, Y., Xiao, T., Li, H., Yi, S., and Wang, X. (2018, January 18\u201322). End-to-end deep kronecker-product matching for person re-identification. Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition, Salt Lake City, UT, USA.","DOI":"10.1109\/CVPR.2018.00720"},{"key":"ref_32","doi-asserted-by":"crossref","first-page":"3254","DOI":"10.1109\/TCSVT.2020.3034981","article-title":"Where to Look and How to Describe: Fashion Image Retrieval with an Attentional Heterogeneous Bilinear Network","volume":"31","author":"Su","year":"2020","journal-title":"IEEE Trans. Circuits Syst. Video Technol."},{"key":"ref_33","doi-asserted-by":"crossref","unstructured":"Verma, S., An, S., Arora, C., and Rai, A. (2018, January 7\u201310). Diversity in fashion recommendation using semantic parsing. Proceedings of the 2018 25th IEEE International Conference on Image Processing (ICIP), Athens, Greece.","DOI":"10.1109\/ICIP.2018.8451164"},{"key":"ref_34","doi-asserted-by":"crossref","unstructured":"Lasserre, J., Rasch, K., and Vollgraf, R. (2018). Studio2shop: From studio photo shoots to fashion articles. arXiv.","DOI":"10.5220\/0006544500370048"}],"container-title":["Sensors"],"original-title":[],"language":"en","link":[{"URL":"https:\/\/www.mdpi.com\/1424-8220\/22\/7\/2660\/pdf","content-type":"unspecified","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2025,10,10]],"date-time":"2025-10-10T22:46:38Z","timestamp":1760136398000},"score":1,"resource":{"primary":{"URL":"https:\/\/www.mdpi.com\/1424-8220\/22\/7\/2660"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2022,3,30]]},"references-count":34,"journal-issue":{"issue":"7","published-online":{"date-parts":[[2022,4]]}},"alternative-id":["s22072660"],"URL":"https:\/\/doi.org\/10.3390\/s22072660","relation":{},"ISSN":["1424-8220"],"issn-type":[{"value":"1424-8220","type":"electronic"}],"subject":[],"published":{"date-parts":[[2022,3,30]]}}}