{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2026,7,4]],"date-time":"2026-07-04T08:19:10Z","timestamp":1783153150269,"version":"3.54.6"},"publisher-location":"New York, NY, USA","reference-count":43,"publisher":"ACM","funder":[{"name":"National Research Foundation of Korea","award":["RS-2024-00406985"],"award-info":[{"award-number":["RS-2024-00406985"]}]},{"name":"Institute of Information & Communications Technology Planning & Evaluation","award":["RS-2022-II220871 & RS-2024-00457882 & RS-2019-II190075"],"award-info":[{"award-number":["RS-2022-II220871 & RS-2024-00457882 & RS-2019-II190075"]}]}],"content-domain":{"domain":["dl.acm.org"],"crossmark-restriction":true},"short-container-title":[],"published-print":{"date-parts":[[2026,4,13]]},"DOI":"10.1145\/3774904.3792248","type":"proceedings-article","created":{"date-parts":[[2026,4,27]],"date-time":"2026-04-27T13:28:36Z","timestamp":1777296516000},"page":"6079-6090","update-policy":"https:\/\/doi.org\/10.1145\/crossmark-policy","source":"Crossref","is-referenced-by-count":0,"title":["Personalized Parameter-Efficient Fine-Tuning of Foundation Models for Multimodal Recommendation"],"prefix":"10.1145","author":[{"ORCID":"https:\/\/orcid.org\/0009-0006-6002-169X","authenticated-orcid":false,"given":"Sunwoo","family":"Kim","sequence":"first","affiliation":[{"name":"KAIST, Seoul, Republic of Korea"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"ORCID":"https:\/\/orcid.org\/0009-0005-5321-2423","authenticated-orcid":false,"given":"Hyunjin","family":"Hwang","sequence":"additional","affiliation":[{"name":"KAIST, Seoul, Republic of Korea"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"ORCID":"https:\/\/orcid.org\/0000-0002-2872-1526","authenticated-orcid":false,"given":"Kijung","family":"Shin","sequence":"additional","affiliation":[{"name":"KAIST, Seoul, Republic of Korea"}],"role":[{"vocabulary":"crossref","role":"author"}]}],"member":"320","published-online":{"date-parts":[[2026,4,12]]},"reference":[{"key":"e_1_3_2_1_1_1","unstructured":"David Arthur and Sergei Vassilvitskii. 2007. k-means the advantages of careful seeding. In SODA."},{"key":"e_1_3_2_1_2_1","doi-asserted-by":"crossref","unstructured":"Matteo Attimonelli Danilo Danese Daniele Malitesta Claudio Pomo Giuseppe Gassi and Tommaso Di Noia. 2024. Ducho 2.0: Towards a more up-to-date unified framework for the extraction of multimodal features in recommendation. In WWW.","DOI":"10.1145\/3589335.3651440"},{"key":"e_1_3_2_1_3_1","doi-asserted-by":"crossref","unstructured":"Zheyu Chen Jinfeng Xu Hewei Wang Shuo Yang Zitong Wan and Haibo Hu. 2025. Hypercomplex Prompt-aware Multimodal Recommendation. In CIKM.","DOI":"10.1145\/3746252.3761174"},{"key":"e_1_3_2_1_4_1","volume-title":"Jose","author":"Fu Junchen","year":"2024","unstructured":"Junchen Fu, Xuri Ge, Xin Xin, Alexandros Karatzoglou, Ioannis Arapakis, Jie Wang, and Joemon M. Jose. 2024a. IISAN: Efficiently Adapting Multimodal Representation for Sequential Recommendation with Decoupled PEFT. In SIGIR."},{"key":"e_1_3_2_1_5_1","volume-title":"Efficient and effective adaptation of multimodal foundation models in sequential recommendation. arXiv preprint arXiv:2411.02992","author":"Fu Junchen","year":"2024","unstructured":"Junchen Fu, Xuri Ge, Xin Xin, Alexandros Karatzoglou, Ioannis Arapakis, Kaiwen Zheng, Yongxin Ni, and Joemon M Jose. 2024b. Efficient and effective adaptation of multimodal foundation models in sequential recommendation. arXiv preprint arXiv:2411.02992 (2024)."},{"key":"e_1_3_2_1_6_1","unstructured":"Junchen Fu Fajie Yuan Yu Song Zheng Yuan Mingyue Cheng Shenghui Cheng Jiaqi Zhang Jie Wang and Yunzhu Pan. 2024c. Exploring adapter-based transfer learning for recommender systems: Empirical studies and practical insights. In WSDM."},{"key":"e_1_3_2_1_7_1","volume-title":"Parameter-Efficient Fine-Tuning for Large Models: A Comprehensive Survey. arXiv preprint arXiv:2403.14608","author":"Han Zeyu","year":"2024","unstructured":"Zeyu Han, Chao Gao, Jinyang Liu, Jeff (Jun) Zhang, and Sai Qian Zhang. 2024. Parameter-Efficient Fine-Tuning for Large Models: A Comprehensive Survey. arXiv preprint arXiv:2403.14608 (2024). https:\/\/arxiv.org\/abs\/2403.14608"},{"key":"e_1_3_2_1_8_1","doi-asserted-by":"publisher","DOI":"10.1145\/3397271.3401063"},{"key":"e_1_3_2_1_9_1","volume-title":"Bridging language and items for retrieval and recommendation. arXiv preprint arXiv:2403.03952","author":"Hou Yupeng","year":"2024","unstructured":"Yupeng Hou, Jiacheng Li, Zhankui He, An Yan, Xiusi Chen, and Julian McAuley. 2024. Bridging language and items for retrieval and recommendation. arXiv preprint arXiv:2403.03952 (2024)."},{"key":"e_1_3_2_1_10_1","unstructured":"Edward J. Hu Yelong Shen Phillip Wallis Zeyuan Allen-Zhu Yuanzhi Li Shean Wang Lu Wang and Weizhu Chen. 2022. LoRA: Low-Rank Adaptation of Large Language Models. In ICLR."},{"key":"e_1_3_2_1_11_1","doi-asserted-by":"crossref","unstructured":"Wang-Cheng Kang and Julian McAuley. 2018. Self-attentive sequential recommendation. In ICDM.","DOI":"10.1109\/ICDM.2018.00035"},{"key":"e_1_3_2_1_12_1","unstructured":"Kyungho Kim Sunwoo Kim Geon Lee and Kijung Shin. 2024. Towards Better Utilization of Multiple Views for Bundle Recommendation. In CIKM."},{"key":"e_1_3_2_1_13_1","volume-title":"ItemRAG: Item-Based Retrieval-Augmented Generation for LLM-Based Recommendation. arXiv preprint arXiv:2511.15141","author":"Kim Sunwoo","year":"2025","unstructured":"Sunwoo Kim, Geon Lee, Kyungho Kim, Jaemin Yoo, and Kijung Shin. 2025. ItemRAG: Item-Based Retrieval-Augmented Generation for LLM-Based Recommendation. arXiv preprint arXiv:2511.15141 (2025)."},{"key":"e_1_3_2_1_14_1","unstructured":"Haokun Liu Derek Tam Muqeeth Mohammed Jay Mohta Tenghao Huang Mohit Bansal and Colin Raffel. 2022. Few-Shot Parameter-Efficient Fine-Tuning is Better and Cheaper than In-Context Learning. In NeurIPS."},{"key":"e_1_3_2_1_15_1","first-page":"1","article-title":"Multimodal recommender systems: A survey","volume":"57","author":"Liu Qidong","year":"2024","unstructured":"Qidong Liu, Jiaxi Hu, Yutian Xiao, Xiangyu Zhao, Jingtong Gao, Wanyu Wang, Qing Li, and Jiliang Tang. 2024a. Multimodal recommender systems: A survey. Comput. Surveys, Vol. 57, 2 (2024), 1-17.","journal-title":"Comput. Surveys"},{"key":"e_1_3_2_1_16_1","unstructured":"Qidong Liu Xian Wu Xiangyu Zhao Yuanshao Zhu Derong Xu Feng Tian and Yefeng Zheng. 2024b. When moe meets llms: Parameter efficient fine-tuning for multi-task medical applications. In SIGIR."},{"key":"e_1_3_2_1_17_1","unstructured":"Shiqin Liu Chaozhuo Li Minjun Zhao Litian Zhang and Jiajun Bu. 2025b. ModalSync: Synchronizing User Behavior with Multimodal Features for Multimodal Pre-training Recommendation. In WWW."},{"key":"e_1_3_2_1_18_1","unstructured":"Weiming Liu Chaochao Chen Jiahe Xu Xinting Liao Fan Wang Xiaolin Zheng Zhihui Fu Ruiguang Pei and Jun Wang. 2025a. Joint similarity item exploration and overlapped user guidance for multi-modal cross-domain recommendation. In WWW."},{"key":"e_1_3_2_1_19_1","unstructured":"Ilya Loshchilov and Frank Hutter. 2019. Decoupled weight decay regularization. In ICLR."},{"key":"e_1_3_2_1_20_1","unstructured":"Xinyu Ma Jiafeng Guo Ruqing Zhang Yixing Fan and Xueqi Cheng. 2022. Scattered or connected? an optimized parameter-efficient tuning approach for information retrieval. In CIKM."},{"key":"e_1_3_2_1_21_1","volume-title":"Sheng Cheng, Changhoon Kim, Tejas Gokhale, Chitta Baral, et al.","author":"Patel Maitreya","year":"2024","unstructured":"Maitreya Patel, Naga Sai Abhiram Kusumba, Sheng Cheng, Changhoon Kim, Tejas Gokhale, Chitta Baral, et al., 2024. Tripletclip: Improving compositional reasoning of clip via synthetic vision-language negatives. In NeurIPS."},{"key":"e_1_3_2_1_22_1","doi-asserted-by":"crossref","unstructured":"Claudio Pomo Matteo Attimonelli Danilo Danese Fedelucio Narducci and Tommaso Di Noia. 2025. Do Recommender Systems Really Leverage Multimodal Content? A Comprehensive Analysis on Multimodal Representations for Recommendation. In CIKM.","DOI":"10.1145\/3746252.3761398"},{"key":"e_1_3_2_1_23_1","doi-asserted-by":"crossref","unstructured":"Filip Radenovic Abhimanyu Dubey Abhishek Kadian Todor Mihaylov Simon Vandenhende Yash Patel Yi Wen Vignesh Ramanathan and Dhruv Mahajan. 2023. Filtering distillation and hard negatives for vision-language pre-training. In CVPR.","DOI":"10.1109\/CVPR52729.2023.00673"},{"key":"e_1_3_2_1_24_1","volume-title":"Chris Hallacy, Aditya Ramesh, Gabriel Goh, Sandhini Agarwal, Girish Sastry, Amanda Askell, Pamela Mishkin, Jack Clark, et al.","author":"Radford Alec","year":"2021","unstructured":"Alec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh, Gabriel Goh, Sandhini Agarwal, Girish Sastry, Amanda Askell, Pamela Mishkin, Jack Clark, et al., 2021. Learning transferable visual models from natural language supervision. In ICML."},{"key":"e_1_3_2_1_25_1","unstructured":"Joshua Robinson Ching-Yao Chuang Suvrit Sra and Stefanie Jegelka. 2021. Contrastive learning with hard negative samples. In ICLR."},{"key":"e_1_3_2_1_26_1","volume-title":"Grad-cam: Visual explanations from deep networks via gradient-based localization. In ICCV.","author":"Selvaraju Ramprasaath R","year":"2017","unstructured":"Ramprasaath R Selvaraju, Michael Cogswell, Abhishek Das, Ramakrishna Vedantam, Devi Parikh, and Dhruv Batra. 2017. Grad-cam: Visual explanations from deep networks via gradient-based localization. In ICCV."},{"key":"e_1_3_2_1_27_1","doi-asserted-by":"crossref","unstructured":"Fei Sun Jun Liu Jian Wu Changhua Pei Xiao Lin Wenwu Ou and Peng Jiang. 2019. BERT4Rec: Sequential recommendation with bidirectional encoder representations from transformer. In CIKM.","DOI":"10.1145\/3357384.3357895"},{"key":"e_1_3_2_1_28_1","doi-asserted-by":"crossref","unstructured":"Shuyao Wang Zhi Zheng Yongduo Sui and Hui Xiong. 2025b. Unleashing the Power of Large Language Model for Denoising Recommendation. In WWW.","DOI":"10.1145\/3696410.3714758"},{"key":"e_1_3_2_1_29_1","volume-title":"Generative Recommendation: Towards Personalized Multimodal Content Generation. In WWW.","author":"Wang Wenjie","year":"2025","unstructured":"Wenjie Wang, Xinyu Lin, Fuli Feng, Xiangnan He, and Tat-Seng Chua. 2025a. Generative Recommendation: Towards Personalized Multimodal Content Generation. In WWW."},{"key":"e_1_3_2_1_30_1","volume-title":"Promptmm: Multi-modal knowledge distillation for recommendation with prompt-tuning. In WWW.","author":"Wei Wei","year":"2024","unstructured":"Wei Wei, Jiabin Tang, Lianghao Xia, Yangqin Jiang, and Chao Huang. 2024. Promptmm: Multi-modal knowledge distillation for recommendation with prompt-tuning. In WWW."},{"key":"e_1_3_2_1_31_1","unstructured":"Liwei Wu Shuqing Li Cho-Jui Hsieh and James Sharpnack. 2020. SSE-PT: Sequential recommendation via personalized transformer. In RecSys."},{"key":"e_1_3_2_1_32_1","volume-title":"Deep multimodal learning with missing modality: A survey. arXiv preprint arXiv:2409.07825","author":"Wu Renjie","year":"2024","unstructured":"Renjie Wu, Hu Wang, Hsiang-Ting Chen, and Gustavo Carneiro. 2024a. Deep multimodal learning with missing modality: A survey. arXiv preprint arXiv:2409.07825 (2024)."},{"key":"e_1_3_2_1_33_1","doi-asserted-by":"publisher","DOI":"10.1109\/TKDE.2024.3357498"},{"key":"e_1_3_2_1_34_1","volume-title":"Mdvt: Enhancing multimodal recommendation with model-agnostic multimodal-driven virtual triplets. In KDD.","author":"Xu Jinfeng","year":"2025","unstructured":"Jinfeng Xu, Zheyu Chen, Jinze Li, Shuo Yang, Hewei Wang, Yijie Li, Mengran Li, Puzhen Wu, and Edith CH Ngai. 2025a. Mdvt: Enhancing multimodal recommendation with model-agnostic multimodal-driven virtual triplets. In KDD."},{"key":"e_1_3_2_1_35_1","volume-title":"Cohesion: Composite graph convolutional network with dual-stage fusion for multimodal recommendation. In SIGIR.","author":"Xu Jinfeng","year":"2025","unstructured":"Jinfeng Xu, Zheyu Chen, Wei Wang, Xiping Hu, Sang-Wook Kim, and Edith CH Ngai. 2025b. Cohesion: Composite graph convolutional network with dual-stage fusion for multimodal recommendation. In SIGIR."},{"key":"e_1_3_2_1_36_1","doi-asserted-by":"publisher","DOI":"10.1145\/3718099"},{"key":"e_1_3_2_1_37_1","doi-asserted-by":"crossref","unstructured":"Zheng Yuan Fajie Yuan Yu Song Youhua Li Junchen Fu Fei Yang Yunzhu Pan and Yongxin Ni. 2023. Where to go next for recommender systems? id-vs. modality-based recommender models revisited. In SIGIR.","DOI":"10.1145\/3539618.3591932"},{"key":"e_1_3_2_1_38_1","doi-asserted-by":"crossref","unstructured":"Jinghao Zhang Yanqiao Zhu Qiang Liu Shu Wu Shuhui Wang and Liang Wang. 2021. Mining latent structures for multimedia recommendation. In MM.","DOI":"10.1145\/3474085.3475259"},{"key":"e_1_3_2_1_39_1","doi-asserted-by":"crossref","first-page":"1","DOI":"10.1145\/3682075","article-title":"Multimodal pre-training for sequential recommendation via contrastive learning","volume":"3","author":"Zhang Lingzi","year":"2024","unstructured":"Lingzi Zhang, Xin Zhou, Zhiwei Zeng, and Zhiqi Shen. 2024. Multimodal pre-training for sequential recommendation via contrastive learning. ACM Transactions on Recommender Systems, Vol. 3, 1 (2024), 1-23.","journal-title":"ACM Transactions on Recommender Systems"},{"key":"e_1_3_2_1_40_1","doi-asserted-by":"crossref","unstructured":"Shengzhe Zhang Liyi Chen Dazhong Shen Chao Wang and Hui Xiong. 2025. Hierarchical Time-Aware Mixture of Experts for Multi-Modal Sequential Recommendation. In WWW.","DOI":"10.1145\/3696410.3714676"},{"key":"e_1_3_2_1_41_1","volume-title":"DVIB: Towards Robust Multimodal Recommender Systems via Variational Information Bottleneck Distillation. In WWW.","author":"Zhao Wenkuan","year":"2025","unstructured":"Wenkuan Zhao, Shanshan Zhong, Yifan Liu, Wushao Wen, Jinghui Qin, Mingfu Liang, and Zhongzhan Huang. 2025. DVIB: Towards Robust Multimodal Recommender Systems via Variational Information Bottleneck Distillation. In WWW."},{"key":"e_1_3_2_1_42_1","doi-asserted-by":"crossref","unstructured":"Xin Zhou and Zhiqi Shen. 2023. A tale of two graphs: Freezing and denoising graph structures for multimodal recommendation. In MM.","DOI":"10.1145\/3581783.3611943"},{"key":"e_1_3_2_1_43_1","doi-asserted-by":"crossref","unstructured":"Xin Zhou Hongyu Zhou Yong Liu Zhiwei Zeng Chunyan Miao Pengwei Wang Yuan You and Feijun Jiang. 2023. Bootstrap latent representations for multi-modal recommendation. In WWW.","DOI":"10.1145\/3543507.3583251"}],"event":{"name":"WWW '26: The ACM Web Conference 2026","location":"Dubai United Arab Emirates","sponsor":["SIGWEB ACM Special Interest Group on Hypertext, Hypermedia, and Web"]},"container-title":["Proceedings of the ACM Web Conference 2026"],"original-title":[],"link":[{"URL":"https:\/\/dl.acm.org\/doi\/pdf\/10.1145\/3774904.3792248","content-type":"unspecified","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2026,7,4]],"date-time":"2026-07-04T07:53:13Z","timestamp":1783151593000},"score":1,"resource":{"primary":{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/3774904.3792248"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2026,4,12]]},"references-count":43,"alternative-id":["10.1145\/3774904.3792248","10.1145\/3774904"],"URL":"https:\/\/doi.org\/10.1145\/3774904.3792248","relation":{},"subject":[],"published":{"date-parts":[[2026,4,12]]},"assertion":[{"value":"2026-04-12","order":3,"name":"published","label":"Published","group":{"name":"publication_history","label":"Publication History"}}]}}