{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2026,7,13]],"date-time":"2026-07-13T18:50:09Z","timestamp":1783968609143,"version":"3.55.0"},"reference-count":67,"publisher":"Association for Computing Machinery (ACM)","issue":"2","license":[{"start":{"date-parts":[[2023,1,25]],"date-time":"2023-01-25T00:00:00Z","timestamp":1674604800000},"content-version":"vor","delay-in-days":0,"URL":"https:\/\/www.acm.org\/publications\/policies\/copyright_policy#Background"}],"funder":[{"name":"Australian Research Council Discovery Project","award":["DP190102353, CE200100025"],"award-info":[{"award-number":["DP190102353, CE200100025"]}]},{"DOI":"10.13039\/501100004543","name":"China Scholarship Council","doi-asserted-by":"crossref","id":[{"id":"10.13039\/501100004543","id-type":"DOI","asserted-by":"crossref"}]}],"content-domain":{"domain":["dl.acm.org"],"crossmark-restriction":true},"short-container-title":["ACM Trans. Inf. Syst."],"published-print":{"date-parts":[[2023,4,30]]},"abstract":"<jats:p>Deep cross-modal retrieval techniques have recently achieved remarkable performance, which also poses severe threats to data privacy potentially. Nowadays, enormous user-generated contents that convey personal information are released and shared on the Internet. One may abuse a retrieval system to pinpoint sensitive information of a particular Internet user, causing privacy leakage. In this article, we propose a data-centric Proactive Privacy-preserving Cross-modal Learning algorithm that fulfills the protection purpose by employing a generator to transform original data into adversarial data with quasi-imperceptible perturbations before releasing them. When the data source is infiltrated, the inside adversarial data can confuse retrieval models under the attacker\u2019s control to make erroneous predictions. We consider the protection under a realistic and challenging setting where the prior knowledge of malicious models is agnostic. To handle this, a surrogate retrieval model is instead introduced, acting as the target to fool. The whole network is trained under a game-theoretical framework, where the generator and the retrieval model persistently evolve to fight against each other. To facilitate the optimization, a Gradient Reversal Layer module is inserted between two models, enabling a one-step learning fashion. Extensive experiments on widely used realistic datasets prove the effectiveness of the proposed method.<\/jats:p>","DOI":"10.1145\/3545799","type":"journal-article","created":{"date-parts":[[2022,6,28]],"date-time":"2022-06-28T13:16:17Z","timestamp":1656422177000},"page":"1-23","update-policy":"https:\/\/doi.org\/10.1145\/crossmark-policy","source":"Crossref","is-referenced-by-count":26,"title":["Proactive Privacy-preserving Learning for Cross-modal Retrieval"],"prefix":"10.1145","volume":"41","author":[{"ORCID":"https:\/\/orcid.org\/0000-0002-6790-2098","authenticated-orcid":false,"given":"Peng-Fei","family":"Zhang","sequence":"first","affiliation":[{"name":"The University of Queensland, Brisbane, QLD, Australia"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"ORCID":"https:\/\/orcid.org\/0000-0002-6390-9890","authenticated-orcid":false,"given":"Guangdong","family":"Bai","sequence":"additional","affiliation":[{"name":"The University of Queensland, Brisbane, QLD, Australia"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"ORCID":"https:\/\/orcid.org\/0000-0003-1395-261X","authenticated-orcid":false,"given":"Hongzhi","family":"Yin","sequence":"additional","affiliation":[{"name":"The University of Queensland, Brisbane, QLD, Australia"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"ORCID":"https:\/\/orcid.org\/0000-0002-9738-4949","authenticated-orcid":false,"given":"Zi","family":"Huang","sequence":"additional","affiliation":[{"name":"The University of Queensland, Brisbane, QLD, Australia"}],"role":[{"vocabulary":"crossref","role":"author"}]}],"member":"320","published-online":{"date-parts":[[2023,1,25]]},"reference":[{"key":"e_1_3_1_2_2","first-page":"1247","volume-title":"Proceedings of the International Conference on Machine Learning","author":"Andrew Galen","year":"2013","unstructured":"Galen Andrew, Raman Arora, Jeff Bilmes, and Karen Livescu. 2013. Deep canonical correlation analysis. In Proceedings of the International Conference on Machine Learning. 1247\u20131255."},{"key":"e_1_3_1_3_2","first-page":"387","volume-title":"Proceedings of the Joint European Conference on Machine Learning and Knowledge Discovery in Databases","author":"Biggio Battista","year":"2013","unstructured":"Battista Biggio, Igino Corona, Davide Maiorca, Blaine Nelson, Nedim \u0160rndi\u0107, Pavel Laskov, Giorgio Giacinto, and Fabio Roli. 2013. Evasion attacks against machine learning at test time. In Proceedings of the Joint European Conference on Machine Learning and Knowledge Discovery in Databases. 387\u2013402."},{"key":"e_1_3_1_4_2","first-page":"10791","volume-title":"Proceedings of the International Conference on Neural Information Processing Systems","author":"Chao Li","year":"2019","unstructured":"Li Chao, Gao Shangqian, Deng Cheng, Xie De, and Liu Wei. 2019. Cross-modal learning with adversarial samples. In Proceedings of the International Conference on Neural Information Processing Systems. 10791\u201310801."},{"key":"e_1_3_1_5_2","volume-title":"Proceedings of the IEEE International Conference on Computer Vision","author":"Chen Zhi","year":"2021","unstructured":"Zhi Chen, Yadan Luo, Ruihong Qiu, Sen Wang, Zi Huang, Jingjing Li, and Zheng Zhang. 2021. Semantics disentangling for generalized zero-shot learning. In Proceedings of the IEEE International Conference on Computer Vision. 8692\u20138700."},{"key":"e_1_3_1_6_2","first-page":"3413","volume-title":"Proceedings of the ACM International Conference on Multimedia","author":"Chen Zhi","year":"2020","unstructured":"Zhi Chen, Sen Wang, Jingjing Li, and Zi Huang. 2020. Rethinking generative zero-shot learning: An ensemble learning perspective for recognising visual patches. In Proceedings of the ACM International Conference on Multimedia. 3413\u20133421."},{"issue":"3","key":"e_1_3_1_7_2","doi-asserted-by":"crossref","first-page":"1","DOI":"10.1145\/3389547","article-title":"Robust unsupervised cross-modal hashing for multimedia retrieval","volume":"38","author":"Cheng Miaomiao","year":"2020","unstructured":"Miaomiao Cheng, Liping Jing, and Michael K. Ng. 2020. Robust unsupervised cross-modal hashing for multimedia retrieval. ACM Trans. Inf. Syst. 38, 3 (2020), 1\u201325.","journal-title":"ACM Trans. Inf. Syst."},{"key":"e_1_3_1_8_2","unstructured":"Valeriia Cherepanova Micah Goldblum Harrison Foley Shiyuan Duan John Dickerson Gavin Taylor and Tom Goldstein. 2021. LowKey: Leveraging adversarial attacks to protect social media users from facial recognition. In Proceedings of the International Conference on Learning Representations ."},{"key":"e_1_3_1_9_2","first-page":"1","volume-title":"Proceedings of the ACM International Conference on Multimedia Information Retrieval","author":"Chua Tat-Seng","year":"2009","unstructured":"Tat-Seng Chua, Jinhui Tang, Richang Hong, Haojie Li, Zhiping Luo, and Yantao Zheng. 2009. NUS-WIDE: A real-world web image database from national university of singapore. In Proceedings of the ACM International Conference on Multimedia Information Retrieval. 1\u20139."},{"key":"e_1_3_1_10_2","doi-asserted-by":"crossref","first-page":"1271","DOI":"10.1109\/TIP.2019.2940693","article-title":"Scalable deep hashing for large-scale social image retrieval","volume":"29","author":"Cui Hui","year":"2019","unstructured":"Hui Cui, Lei Zhu, Jingjing Li, Yang Yang, and Liqiang Nie. 2019. Scalable deep hashing for large-scale social image retrieval. IEEE Trans. Image Process. 29 (2019), 1271\u20131284.","journal-title":"IEEE Trans. Image Process."},{"key":"e_1_3_1_11_2","first-page":"2075","volume-title":"Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition","author":"Ding Guiguang","year":"2014","unstructured":"Guiguang Ding, Yuchen Guo, and Jile Zhou. 2014. Collective matrix factorization hashing for multimodal data. In Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition. 2075\u20132082."},{"key":"e_1_3_1_12_2","first-page":"9185","volume-title":"Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition","author":"Dong Yinpeng","year":"2018","unstructured":"Yinpeng Dong, Fangzhou Liao, Tianyu Pang, Hang Su, Jun Zhu, Xiaolin Hu, and Jianguo Li. 2018. Boosting adversarial attacks with momentum. In Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition. 9185\u20139193."},{"key":"e_1_3_1_13_2","first-page":"1180","volume-title":"Proceedings of the International Conference on Machine Learning","author":"Ganin Yaroslav","year":"2015","unstructured":"Yaroslav Ganin and Victor Lempitsky. 2015. Unsupervised domain adaptation by backpropagation. In Proceedings of the International Conference on Machine Learning. 1180\u20131189."},{"issue":"1","key":"e_1_3_1_14_2","article-title":"Domain-adversarial training of neural networks","volume":"17","author":"Ganin Yaroslav","year":"2016","unstructured":"Yaroslav Ganin, Evgeniya Ustinova, Hana Ajakan, Pascal Germain, Hugo Larochelle, Fran\u00e7ois Laviolette, Mario Marchand, and Victor Lempitsky. 2016. Domain-adversarial training of neural networks. J. Mach. Learn. Res. 17, 1 (2016), 2096\u20132030.","journal-title":"J. Mach. Learn. Res."},{"key":"e_1_3_1_15_2","unstructured":"Ian J. Goodfellow Jonathon Shlens and Christian Szegedy. 2014. Explaining and harnessing adversarial examples. In Proceedings of the International Conference on Learning Representations ."},{"issue":"1","key":"e_1_3_1_16_2","first-page":"723","article-title":"A kernel two-sample test","volume":"13","author":"Gretton Arthur","year":"2012","unstructured":"Arthur Gretton, Karsten M. Borgwardt, Malte J. Rasch, Bernhard Sch\u00f6lkopf, and Alexander Smola. 2012. A kernel two-sample test. J. Mach. Learn. Res. 13, 1 (2012), 723\u2013773.","journal-title":"J. Mach. Learn. Res."},{"key":"e_1_3_1_17_2","article-title":"Badnets: Identifying vulnerabilities in the machine learning model supply chain","author":"Gu Tianyu","year":"2017","unstructured":"Tianyu Gu, Brendan Dolan-Gavitt, and Siddharth Garg. 2017. Badnets: Identifying vulnerabilities in the machine learning model supply chain. arXiv:1708.06733. Retrieved from https:\/\/arxiv.org\/abs\/1708.06733.","journal-title":"arXiv:1708.06733"},{"key":"e_1_3_1_18_2","first-page":"770","volume-title":"Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition","author":"He Kaiming","year":"2016","unstructured":"Kaiming He, Xiangyu Zhang, Shaoqing Ren, and Jian Sun. 2016. Deep residual learning for image recognition. In Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition. 770\u2013778."},{"key":"e_1_3_1_19_2","doi-asserted-by":"crossref","first-page":"162","DOI":"10.1007\/978-1-4612-4380-9_14","volume-title":"Breakthroughs in Statistics","author":"Hotelling Harold","year":"1992","unstructured":"Harold Hotelling. 1992. Relations between two sets of variates. In Breakthroughs in Statistics. 162\u2013190."},{"key":"e_1_3_1_20_2","first-page":"3123","volume-title":"Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition","author":"Hu Hengtong","year":"2020","unstructured":"Hengtong Hu, Lingxi Xie, Richang Hong, and Qi Tian. 2020. Creating something from nothing: Unsupervised knowledge distillation for cross-modal hashing. In Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition. 3123\u20133132."},{"issue":"6","key":"e_1_3_1_21_2","first-page":"2770","article-title":"Collective reconstructive embeddings for cross-modal hashing","volume":"28","author":"Hu Mengqiu","year":"2018","unstructured":"Mengqiu Hu, Yang Yang, Fumin Shen, Ning Xie, Richang Hong, and Heng Tao Shen. 2018. Collective reconstructive embeddings for cross-modal hashing. IEEE Trans. Image Process. 28, 6 (2018), 2770\u20132784.","journal-title":"IEEE Trans. Image Process."},{"key":"e_1_3_1_22_2","doi-asserted-by":"crossref","first-page":"39","DOI":"10.1145\/1460096.1460104","volume-title":"Proceedings of the ACM International Conference on Multimedia Information Retrieval","author":"Huiskes Mark J.","year":"2008","unstructured":"Mark J. Huiskes and Michael S. Lew. 2008. The MIR flickr retrieval evaluation. In Proceedings of the ACM International Conference on Multimedia Information Retrieval. 39\u201343."},{"key":"e_1_3_1_23_2","first-page":"125","volume-title":"Proceedings of the International Conference in Neural Information Processing Systems","author":"Ilyas Andrew","year":"2019","unstructured":"Andrew Ilyas, Shibani Santurkar, Dimitris Tsipras, Logan Engstrom, Brandon Tran, and Aleksander Madry. 2019. Adversarial examples are not bugs, they are features. In Proceedings of the International Conference in Neural Information Processing Systems. 125\u2013136."},{"key":"e_1_3_1_24_2","first-page":"3232","volume-title":"Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition","author":"Jiang Qing-Yuan","year":"2017","unstructured":"Qing-Yuan Jiang and Wu-Jun Li. 2017. Deep cross-modal hashing. In Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition. 3232\u20133240."},{"issue":"1","key":"e_1_3_1_25_2","first-page":"188","article-title":"Multi-view discriminant analysis","volume":"38","author":"Kan Meina","year":"2015","unstructured":"Meina Kan, Shiguang Shan, Haihong Zhang, Shihong Lao, and Xilin Chen. 2015. Multi-view discriminant analysis. IEEE Trans. Pattern Anal. Mach. Intell. 38, 1 (2015), 188\u2013194.","journal-title":"IEEE Trans. Pattern Anal. Mach. Intell."},{"key":"e_1_3_1_26_2","first-page":"4893","volume-title":"Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition","author":"Kang Guoliang","year":"2019","unstructured":"Guoliang Kang, Lu Jiang, Yi Yang, and Alexander G. Hauptmann. 2019. Contrastive adaptation network for unsupervised domain adaptation. In Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition. 4893\u20134902."},{"key":"e_1_3_1_27_2","doi-asserted-by":"crossref","unstructured":"Yoon Kim. 2014. Convolutional neural networks for sentence classification. In Proceedings of the Conference on Empirical Methods in Natural Language Processing . 1746\u20131751.","DOI":"10.3115\/v1\/D14-1181"},{"key":"e_1_3_1_28_2","first-page":"1360","volume-title":"Proceedings of the International Joint Conference on Artificial Intelligence","author":"Kumar Shaishav","year":"2011","unstructured":"Shaishav Kumar and Raghavendra Udupa. 2011. Learning hash functions for cross-view similarity search. In Proceedings of the International Joint Conference on Artificial Intelligence. 1360\u20131365."},{"key":"e_1_3_1_29_2","first-page":"421","volume-title":"Proceedings of the ACM SIGKDD International Conference on Knowledge Discovery and Data Mining","author":"Li Chao","year":"2020","unstructured":"Chao Li, Haoteng Tang, Cheng Deng, Liang Zhan, and Wei Liu. 2020. Vulnerability vs. reliability: Disentangled adversarial examples for cross-modal learning. In Proceedings of the ACM SIGKDD International Conference on Knowledge Discovery and Data Mining. 421\u2013429."},{"key":"e_1_3_1_30_2","volume-title":"Proceedings of the International Conference on Neural Information Processing Systems","author":"Li Qizhang","year":"2020","unstructured":"Qizhang Li, Yiwen Guo, and Hao Chen. 2020. Practical no-box adversarial attacks against DNNs. In Proceedings of the International Conference on Neural Information Processing Systems. 12849\u201312860."},{"key":"e_1_3_1_31_2","first-page":"1379","volume-title":"Proceedings of the ACM SIGIR International Conference on Research and Development in Information Retrieval","author":"Liu Song","year":"2020","unstructured":"Song Liu, Shengsheng Qian, Yang Guan, Jiawei Zhan, and Long Ying. 2020. Joint-modal distribution-based similarity hashing for large-scale unsupervised deep cross-modal retrieval. In Proceedings of the ACM SIGIR International Conference on Research and Development in Information Retrieval. 1379\u20131388."},{"key":"e_1_3_1_32_2","first-page":"1107","volume-title":"Proceedings of the IEEE International Conference on Computer Vision","author":"Liu Xianglong","year":"2015","unstructured":"Xianglong Liu, Lei Huang, Cheng Deng, Jiwen Lu, and Bo Lang. 2015. Multi-view complementary hash tables for nearest neighbor search. In Proceedings of the IEEE International Conference on Computer Vision. 1107\u20131115."},{"key":"e_1_3_1_33_2","first-page":"715","volume-title":"Proceedings of the ACM SIGIR International Conference on Research and Development in Information Retrieval","author":"Lu Xu","year":"2019","unstructured":"Xu Lu, Lei Zhu, Zhiyong Cheng, Liqiang Nie, and Huaxiang Zhang. 2019. Online multi-modal hashing with dynamic query-adaption. In Proceedings of the ACM SIGIR International Conference on Research and Development in Information Retrieval. 715\u2013724."},{"key":"e_1_3_1_34_2","first-page":"2341","volume-title":"Proceedings of the ACM International Conference on Multimedia","author":"Luo Yadan","year":"2019","unstructured":"Yadan Luo, Zi Huang, Zheng Zhang, Ziwei Wang, Jingjing Li, and Yang Yang. 2019. Curiosity-driven reinforcement learning for diverse visual paragraph generation. In Proceedings of the ACM International Conference on Multimedia. 2341\u20132350."},{"key":"e_1_3_1_35_2","first-page":"2574","volume-title":"Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition","author":"Moosavi-Dezfooli Seyed-Mohsen","year":"2016","unstructured":"Seyed-Mohsen Moosavi-Dezfooli, Alhussein Fawzi, and Pascal Frossard. 2016. Deepfool: A simple and accurate method to fool deep neural networks. In Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition. 2574\u20132582."},{"key":"e_1_3_1_36_2","doi-asserted-by":"crossref","unstructured":"Konda Reddy Mopuri Utsav Garg and R. Venkatesh Babu. 2017. Fast feature fool: A data independent approach to universal adversarial perturbations. In Proceedings of the British Machine Vision Conference .","DOI":"10.5244\/C.31.30"},{"key":"e_1_3_1_37_2","first-page":"1491","volume-title":"Proceedings of the IEEE International Conference on Computer Vision","author":"Oh Seong Joon","year":"2017","unstructured":"Seong Joon Oh, Mario Fritz, and Bernt Schiele. 2017. Adversarial image perturbation for privacy protection a game theory perspective. In Proceedings of the IEEE International Conference on Computer Vision. 1491\u20131500."},{"key":"e_1_3_1_38_2","article-title":"Transferability in machine learning: From phenomena to black-box attacks using adversarial samples","author":"Papernot Nicolas","year":"2016","unstructured":"Nicolas Papernot, Patrick McDaniel, and Ian Goodfellow. 2016. Transferability in machine learning: From phenomena to black-box attacks using adversarial samples. arXiv:1605.07277. Retrieved from https:\/\/arxiv.org\/abs\/1605.07277.","journal-title":"arXiv:1605.07277"},{"issue":"3","key":"e_1_3_1_39_2","doi-asserted-by":"crossref","first-page":"1","DOI":"10.1145\/3382764","article-title":"Exploiting cross-session information for session-based recommendation with graph neural networks","volume":"38","author":"Qiu Ruihong","year":"2020","unstructured":"Ruihong Qiu, Zi Huang, Jingjing Li, and Hongzhi Yin. 2020. Exploiting cross-session information for session-based recommendation with graph neural networks. ACM Trans. Inf. Syst. 38, 3 (2020), 1\u201323.","journal-title":"ACM Trans. Inf. Syst."},{"key":"e_1_3_1_40_2","first-page":"4094","volume-title":"Proceedings of the IEEE International Conference on Computer Vision","author":"Ranjan Viresh","year":"2015","unstructured":"Viresh Ranjan, Nikhil Rasiwasia, and C. V. Jawahar. 2015. Multi-label cross-modal retrieval. In Proceedings of the IEEE International Conference on Computer Vision. 4094\u20134102."},{"issue":"4","key":"e_1_3_1_41_2","doi-asserted-by":"crossref","first-page":"1","DOI":"10.1145\/3394592","article-title":"CRSAL: Conversational recommender systems with adversarial learning","volume":"38","author":"Ren Xuhui","year":"2020","unstructured":"Xuhui Ren, Hongzhi Yin, Tong Chen, Hao Wang, Nguyen Quoc Viet Hung, Zi Huang, and Xiangliang Zhang. 2020. CRSAL: Conversational recommender systems with adversarial learning. ACM Trans. Inf. Syst. 38, 4 (2020), 1\u201340.","journal-title":"ACM Trans. Inf. Syst."},{"key":"e_1_3_1_42_2","first-page":"2263","article-title":"Equivalence of distance-based and RKHS-based statistics in hypothesis testing","author":"Sejdinovic Dino","year":"2013","unstructured":"Dino Sejdinovic, Bharath Sriperumbudur, Arthur Gretton, and Kenji Fukumizu. 2013. Equivalence of distance-based and RKHS-based statistics in hypothesis testing. Ann. Stat. (2013), 2263\u20132291.","journal-title":"Ann. Stat."},{"key":"e_1_3_1_43_2","first-page":"6103","volume-title":"Proceedings of the International Conference on Neural Information Processing Systems","author":"Shafahi Ali","year":"2018","unstructured":"Ali Shafahi, W. Ronny Huang, Mahyar Najibi, Octavian Suciu, Christoph Studer, Tudor Dumitras, and Tom Goldstein. 2018. Poison frogs! targeted clean-label poisoning attacks on neural networks. In Proceedings of the International Conference on Neural Information Processing Systems. 6103\u20136113."},{"key":"e_1_3_1_44_2","first-page":"1589\u2013 1604","volume-title":"Proceedings of the USENIX Security Symposium","author":"Shan Shawn","year":"2020","unstructured":"Shawn Shan, Emily Wenger, Jiayun Zhang, Huiying Li, Haitao Zheng, and Ben Y. Zhao. 2020. Fawkes: Protecting privacy against unauthorized deep learning models. In Proceedings of the USENIX Security Symposium. 1589\u2013 1604."},{"issue":"10","key":"e_1_3_1_45_2","doi-asserted-by":"crossref","first-page":"3351","DOI":"10.1109\/TKDE.2020.2970050","article-title":"Exploiting subspace relation in semantic labels for cross-modal hashing","volume":"33","author":"Shen Heng Tao","year":"2020","unstructured":"Heng Tao Shen, Luchen Liu, Yang Yang, Xing Xu, Zi Huang, Fumin Shen, and Richang Hong. 2020. Exploiting subspace relation in semantic labels for cross-modal hashing. IEEE Trans. Knowl. Data Eng. 33, 10 (2020), 3351\u20133365.","journal-title":"IEEE Trans. Knowl. Data Eng."},{"key":"e_1_3_1_46_2","article-title":"Very deep convolutional networks for large-scale image recognition","author":"Simonyan Karen","year":"2014","unstructured":"Karen Simonyan and Andrew Zisserman. 2014. Very deep convolutional networks for large-scale image recognition. arXiv:1409.1556. Retrieved from https:\/\/arxiv.org\/abs\/1409.1556.","journal-title":"arXiv:1409.1556"},{"key":"e_1_3_1_47_2","first-page":"5","volume-title":"Proceedings of the ACM SIGIR International Conference on Research and Development in Information Retrieval","author":"Song Xuemeng","year":"2018","unstructured":"Xuemeng Song, Fuli Feng, Xianjing Han, Xin Yang, Wei Liu, and Liqiang Nie. 2018. Neural compatibility modeling with attentive knowledge distillation. In Proceedings of the ACM SIGIR International Conference on Research and Development in Information Retrieval. 5\u201314."},{"key":"e_1_3_1_48_2","first-page":"3027","volume-title":"Proceedings of the IEEE International Conference on Computer Vision","author":"Su Shupeng","year":"2019","unstructured":"Shupeng Su, Zhisheng Zhong, and Chao Zhang. 2019. Deep joint-semantics reconstructing hashing for large-scale unsupervised cross-modal retrieval. In Proceedings of the IEEE International Conference on Computer Vision. 3027\u20133035."},{"key":"e_1_3_1_49_2","first-page":"1","volume-title":"Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition","author":"Szegedy Christian","year":"2015","unstructured":"Christian Szegedy, Wei Liu, Yangqing Jia, Pierre Sermanet, Scott Reed, Dragomir Anguelov, Dumitru Erhan, Vincent Vanhoucke, and Andrew Rabinovich. 2015. Going deeper with convolutions. In Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition. 1\u20139."},{"key":"e_1_3_1_50_2","doi-asserted-by":"crossref","unstructured":"Xing Xu Kaiyi Lin Yang Yang Alan Hanjalic and Heng Tao Shen. 2022. Joint feature synthesis and embedding: Adversarial cross-modal retrieval revisited. IEEE Trans. Pattern Anal. Mach. Intell. 44 6 (2022) 3030\u20133047.","DOI":"10.1109\/TPAMI.2020.3045530"},{"key":"e_1_3_1_51_2","volume-title":"Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition Workshops","author":"Thys Simen","year":"2019","unstructured":"Simen Thys, Wiebe Van Ranst, and Toon Goedem\u00e9. 2019. Fooling automated surveillance cameras: Adversarial patches to attack person detection. In Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition Workshops."},{"key":"e_1_3_1_52_2","article-title":"The space of transferable adversarial examples","author":"Tram\u00e8r Florian","year":"2017","unstructured":"Florian Tram\u00e8r, Nicolas Papernot, Ian Goodfellow, Dan Boneh, and Patrick McDaniel. 2017. The space of transferable adversarial examples. arXiv:1704.03453. Retrieved from https:\/\/arxiv.org\/abs\/1704.03453.","journal-title":"arXiv:1704.03453"},{"key":"e_1_3_1_53_2","first-page":"154","volume-title":"Proceedings of the AAAI Conference on Artificial Intelligence","author":"Wang Bokun","year":"2017","unstructured":"Bokun Wang, Yang Yang, Xing Xu, Alan Hanjalic, and Heng Tao Shen. 2017. Adversarial cross-modal retrieval. In Proceedings of the AAAI Conference on Artificial Intelligence. 154\u2013162."},{"key":"e_1_3_1_54_2","first-page":"1","article-title":"Fast-adapting and privacy-preserving federated recommender system","author":"Wang Qinyong","year":"2021","unstructured":"Qinyong Wang, Hongzhi Yin, Tong Chen, Junliang Yu, Alexander Zhou, and Xiangliang Zhang. 2021. Fast-adapting and privacy-preserving federated recommender system. The VLDB J. (2021), 1\u201320.","journal-title":"The VLDB J."},{"key":"e_1_3_1_55_2","first-page":"1083","volume-title":"Proceedings of the International Conference on Machine Learning","author":"Wang Weiran","year":"2015","unstructured":"Weiran Wang, Raman Arora, Karen Livescu, and Jeff Bilmes. 2015. On deep multi-view representation learning. In Proceedings of the International Conference on Machine Learning. 1083\u20131092."},{"key":"e_1_3_1_56_2","first-page":"2900","volume-title":"Proceedings of the Web Conference","author":"Wang Yongxin","year":"2021","unstructured":"Yongxin Wang, Zhen-Duo Chen, Xin Luo, and Xin-Shun Xu. 2021. High-dimensional sparse cross-modal hashing with fine-grained similarity embedding. In Proceedings of the Web Conference. 2900\u20132909."},{"key":"e_1_3_1_57_2","doi-asserted-by":"crossref","first-page":"3626","DOI":"10.1109\/TIP.2020.2963957","article-title":"Multi-task consistency-preserving adversarial hashing for cross-modal retrieval","volume":"29","author":"Xie De","year":"2020","unstructured":"De Xie, Cheng Deng, Chao Li, Xianglong Liu, and Dacheng Tao. 2020. Multi-task consistency-preserving adversarial hashing for cross-modal retrieval. IEEE Trans. Image Process. 29 (2020), 3626\u20133637.","journal-title":"IEEE Trans. Image Process."},{"key":"e_1_3_1_58_2","article-title":"Joint feature synthesis and embedding: Adversarial cross-modal retrieval revisited","author":"Xu Xing","year":"2020","unstructured":"Xing Xu, Kaiyi Lin, Yang Yang, Alan Hanjalic, and Heng Tao Shen. 2020. Joint feature synthesis and embedding: Adversarial cross-modal retrieval revisited. IEEE Trans. Pattern Anal. Mach. Intell. (2020).","journal-title":"IEEE Trans. Pattern Anal. Mach. Intell."},{"key":"e_1_3_1_59_2","first-page":"3386","volume-title":"Proceedings of the ACM International Conference on Multimedia","author":"Zhan Yu-Wei","year":"2020","unstructured":"Yu-Wei Zhan, Xin Luo, Yongxin Wang, and Xin-Shun Xu. 2020. Supervised hierarchical deep hashing for cross-modal retrieval. In Proceedings of the ACM International Conference on Multimedia. 3386\u20133394."},{"key":"e_1_3_1_60_2","first-page":"7","volume-title":"Proceedings of the AAAI Conference on Artificial Intelligence","author":"Zhang Dongqing","year":"2014","unstructured":"Dongqing Zhang and Wu-Jun Li. 2014. Large-scale supervised multimodal hashing with semantic correlation maximization. In Proceedings of the AAAI Conference on Artificial Intelligence. 7\u201313."},{"key":"e_1_3_1_61_2","first-page":"3369","volume-title":"Proceedings of the AAAI Conference on Artificial Intelligence","author":"Zhang Peng-Fei","year":"2021","unstructured":"Peng-Fei Zhang, Zi Huang, and Xin-Shun Xu. 2021. Proactive privacy-preserving learning for retrieval. In Proceedings of the AAAI Conference on Artificial Intelligence. 3369\u20133376."},{"issue":"2","key":"e_1_3_1_62_2","doi-asserted-by":"crossref","first-page":"563","DOI":"10.1007\/s11280-020-00859-y","article-title":"High-order nonlocal hashing for unsupervised cross-modal retrieval","volume":"24","author":"Zhang Peng-Fei","year":"2021","unstructured":"Peng-Fei Zhang, Yadan Luo, Zi Huang, Xin-Shun Xu, and Jingkuan Song. 2021. High-order nonlocal hashing for unsupervised cross-modal retrieval. World Wide Web 24, 2 (2021), 563\u2013583.","journal-title":"World Wide Web"},{"key":"e_1_3_1_63_2","first-page":"1415","volume-title":"Proceedings of the ACM International Conference on Web Search and Data Mining","author":"Zhang Shijie","year":"2022","unstructured":"Shijie Zhang, Hongzhi Yin, Tong Chen, Zi Huang, Quoc Viet Hung Nguyen, and Lizhen Cui. 2022. Pipattack: Poisoning federated recommender systems for manipulating item promotion. In Proceedings of the ACM International Conference on Web Search and Data Mining. 1415\u20131423."},{"key":"e_1_3_1_64_2","article-title":"Deep multimodal transfer learning for cross-modal retrieval","author":"Zhen Liangli","year":"2022","unstructured":"Liangli Zhen, Peng Hu, Xi Peng, Rick Siow Mong Goh, and Joey Tianyi Zhou. 2022. Deep multimodal transfer learning for cross-modal retrieval. IEEE Trans. Neural Netw. Learn. Syst. 33, 2 (2022), 798\u2013810.","journal-title":"IEEE Trans. Neural Netw. Learn. Syst."},{"key":"e_1_3_1_65_2","first-page":"10394","volume-title":"Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition","author":"Zhen Liangli","year":"2019","unstructured":"Liangli Zhen, Peng Hu, Xu Wang, and Dezhong Peng. 2019. Deep supervised cross-modal retrieval. In Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition. 10394\u201310403."},{"key":"e_1_3_1_66_2","first-page":"1376","volume-title":"Proceedings of the International Conference on Neural Information Processing Systems","author":"Zhen Yi","year":"2012","unstructured":"Yi Zhen and Dit-Yan Yeung. 2012. Co-regularized hashing for multimodal data. In Proceedings of the International Conference on Neural Information Processing Systems. 1376\u20131384."},{"key":"e_1_3_1_67_2","first-page":"143","volume-title":"Proceedings of the ACM International Conference on Multimedia","author":"Zhu Xiaofeng","year":"2013","unstructured":"Xiaofeng Zhu, Zi Huang, Heng Tao Shen, and Xin Zhao. 2013. Linear cross-modal hashing for efficient multimedia search. In Proceedings of the ACM International Conference on Multimedia. 143\u2013152."},{"key":"e_1_3_1_68_2","doi-asserted-by":"crossref","unstructured":"Peng-Fei Zhang Chuan-Xiang Li Meng-Yuan Liu Liqiang Nie and Xin-Shun Xu. 2017. Semi-relaxation supervised hashing for cross-modal retrieval. In Proceedings of the ACM International Conference on Multimedia . 1762\u20131770.","DOI":"10.1145\/3123266.3123320"}],"container-title":["ACM Transactions on Information Systems"],"original-title":[],"language":"en","link":[{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/3545799","content-type":"unspecified","content-version":"vor","intended-application":"text-mining"},{"URL":"https:\/\/dl.acm.org\/doi\/pdf\/10.1145\/3545799","content-type":"unspecified","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2025,6,17]],"date-time":"2025-06-17T19:02:46Z","timestamp":1750186966000},"score":1,"resource":{"primary":{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/3545799"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2023,1,25]]},"references-count":67,"journal-issue":{"issue":"2","published-print":{"date-parts":[[2023,4,30]]}},"alternative-id":["10.1145\/3545799"],"URL":"https:\/\/doi.org\/10.1145\/3545799","relation":{},"ISSN":["1046-8188","1558-2868"],"issn-type":[{"value":"1046-8188","type":"print"},{"value":"1558-2868","type":"electronic"}],"subject":[],"published":{"date-parts":[[2023,1,25]]},"assertion":[{"value":"2021-11-28","order":0,"name":"received","label":"Received","group":{"name":"publication_history","label":"Publication History"}},{"value":"2022-06-05","order":1,"name":"accepted","label":"Accepted","group":{"name":"publication_history","label":"Publication History"}},{"value":"2023-01-25","order":2,"name":"published","label":"Published","group":{"name":"publication_history","label":"Publication History"}}]}}