{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2026,5,27]],"date-time":"2026-05-27T20:00:16Z","timestamp":1779912016507,"version":"3.53.1"},"reference-count":52,"publisher":"Association for Computing Machinery (ACM)","issue":"8","license":[{"start":{"date-parts":[[2024,8,16]],"date-time":"2024-08-16T00:00:00Z","timestamp":1723766400000},"content-version":"vor","delay-in-days":0,"URL":"https:\/\/www.acm.org\/publications\/policies\/copyright_policy#Background"}],"funder":[{"DOI":"10.13039\/501100001809","name":"National Natural Science Foundation of China","doi-asserted-by":"crossref","award":["62272191, 62372211"],"award-info":[{"award-number":["62272191, 62372211"]}],"id":[{"id":"10.13039\/501100001809","id-type":"DOI","asserted-by":"crossref"}]},{"name":"International Science and Technology Cooperation Program of Jilin Province","award":["20230402076GH, 20240402067GH"],"award-info":[{"award-number":["20230402076GH, 20240402067GH"]}]},{"name":"Science and Technology Development Program of Jilin Province","award":["20220201153GX"],"award-info":[{"award-number":["20220201153GX"]}]},{"name":"Fifth Electronics Research Institute of the Ministry of Industry and Information Technology","award":["HK202303528"],"award-info":[{"award-number":["HK202303528"]}]}],"content-domain":{"domain":["dl.acm.org"],"crossmark-restriction":true},"short-container-title":["ACM Trans. Knowl. Discov. Data"],"published-print":{"date-parts":[[2024,9,30]]},"abstract":"<jats:p>Deep learning models are often trained on datasets that are limited in size and distribution, which may not fully represent the entire range of data encountered in practice. Thus, making deep learning models generalize to out-of-distribution data has received a significant amount of attention in recent studies due to the critical importance of this ability in real-world applications. Meta learning as an effective knowledge transfer paradigm, which learns a base model with high generalization ability to adapt to new data distributions by minimizing domain shifts across tasks during meta-training. However, most existing meta learning methods assume that the base model can access the labels of different domains, and this assumption is demanding in many real application scenarios. In addition, these methods focus on narrowing data-level domain shifts, while ignoring task-level domain shifts, which may lead to inadequate or even negative transfer. Inspired by human learners who use induction to learn and master new tasks, we propose a novel domain-aware meta learning framework for out-of-distribution generalization, termed SMLG. This framework enables the base model to generalize effectively to unseen domains without relying on domain-specific labels. Specifically, we develop a domain-aware transformation module to obtain meta representation and pseudo domain labels. As a result, the base model can be trained robustly without the need for direct domain label input. Furthermore, to investigate the impact of domain shifts at different levels, we introduce a joint loss function that combines cross-entropy with a domain alignment constraint. Extensive experiments on benchmark datasets demonstrate the efficacy of our framework.<\/jats:p>","DOI":"10.1145\/3676558","type":"journal-article","created":{"date-parts":[[2024,7,3]],"date-time":"2024-07-03T11:50:42Z","timestamp":1720007442000},"page":"1-24","update-policy":"https:\/\/doi.org\/10.1145\/crossmark-policy","source":"Crossref","is-referenced-by-count":4,"title":["Towards Domain-Aware Stable Meta Learning for Out-of-Distribution Generalization"],"prefix":"10.1145","volume":"18","author":[{"ORCID":"https:\/\/orcid.org\/0000-0002-7199-0836","authenticated-orcid":false,"given":"Mingchen","family":"Sun","sequence":"first","affiliation":[{"name":"The College of Computer Science and Technology, Jilin University, Changchun, China"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"ORCID":"https:\/\/orcid.org\/0000-0002-3575-1395","authenticated-orcid":false,"given":"Yingji","family":"Li","sequence":"additional","affiliation":[{"name":"The College of Computer Science and Technology, Jilin University, Changchun, China"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"ORCID":"https:\/\/orcid.org\/0000-0002-3288-5195","authenticated-orcid":false,"given":"Ying","family":"Wang","sequence":"additional","affiliation":[{"name":"The College of Computer Science and Technology, Jilin University, Changchun, China and Key Laboratory of Symbolic Computation and Knowledge Engineering of Ministry of Education, Jilin University, Changchun, China"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"ORCID":"https:\/\/orcid.org\/0000-0001-9448-7689","authenticated-orcid":false,"given":"Xin","family":"Wang","sequence":"additional","affiliation":[{"name":"The School of Artificial Intelligence, Jilin University, Changchun, China"}],"role":[{"vocabulary":"crossref","role":"author"}]}],"member":"320","published-online":{"date-parts":[[2024,8,16]]},"reference":[{"key":"e_1_3_1_2_2","first-page":"10040","volume-title":"Proceedings of Advances in Neural Information Processing Systems (NeurIPS)","author":"Arik Sercan \u00d6mer","year":"2018","unstructured":"Sercan \u00d6mer Arik, Jitong Chen, Kainan Peng, Wei Ping, and Yanqi Zhou. 2018. Neural Voice Cloning with a Few Samples. In Proceedings of Advances in Neural Information Processing Systems (NeurIPS). 10040\u201310050."},{"key":"e_1_3_1_3_2","unstructured":"Martin Arjovsky L\u00e9on Bottou Ishaan Gulrajani and David Lopez-Paz. 2019. Invariant Risk Minimization. arXiv:1907.02893. Retrieved from http:\/\/arxiv.org\/abs\/1907.02893"},{"key":"e_1_3_1_4_2","first-page":"998","volume-title":"Proceedings of Advances in Neural Information Processing Systems (NeurIPS)","volume":"31","author":"Balaji Yogesh","year":"2018","unstructured":"Yogesh Balaji, Swami Sankaranarayanan, and Rama Chellappa. 2018. Metareg: Towards Domain Generalization Using Meta-Regularization. In Proceedings of Advances in Neural Information Processing Systems (NeurIPS), Vol. 31. 998\u20131008."},{"key":"e_1_3_1_5_2","first-page":"137","volume-title":"Proceedings of Advances in Neural Information Processing Systems (NeurIPS)","author":"Ben-David Shai","year":"2006","unstructured":"Shai Ben-David, John Blitzer, Koby Crammer, and Fernando Pereira. 2006. Analysis of Representations for Domain Adaptation. In Proceedings of Advances in Neural Information Processing Systems (NeurIPS). 137\u2013144."},{"key":"e_1_3_1_6_2","unstructured":"Hakan Bilen and Andrea Vedaldi. 2017. Universal Representations: The Missing Link Between Faces Text Planktons and Cat Breeds. arXiv:1701.07275. Retrieved from http:\/\/arxiv.org\/abs\/1701.07275"},{"key":"e_1_3_1_7_2","doi-asserted-by":"publisher","DOI":"10.24963\/ijcai.2022\/758"},{"key":"e_1_3_1_8_2","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR.2019.00233"},{"key":"e_1_3_1_9_2","doi-asserted-by":"publisher","DOI":"10.1109\/ICCV48922.2021.01576"},{"key":"e_1_3_1_10_2","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR52688.2022.00698"},{"key":"e_1_3_1_11_2","first-page":"396","volume-title":"Proceedings of the Advances in Neural Information Processing Systems 33: Annual Conference on Neural Information Processing Systems (NeurIPS)","author":"Chen Jiaxin","year":"2020","unstructured":"Jiaxin Chen, Xiao-Ming Wu, Yanke Li, Qimai Li, Li-Ming Zhan, and Fu-Lai Chung. 2020. A Closer Look at the Training Strategy for Modern Meta-Learning. In Proceedings of the Advances in Neural Information Processing Systems 33: Annual Conference on Neural Information Processing Systems (NeurIPS). 396\u2013406."},{"key":"e_1_3_1_12_2","first-page":"418","volume-title":"Neurocomputing","volume":"467","author":"Chen Keyu","year":"2022","unstructured":"Keyu Chen, Di Zhuang, and J. Morris Chang. 2022b. Discriminative Adversarial Domain Generalization with Meta-Learning Based Cross-Domain Validation. Neurocomputing 467 (2022), 418\u2013426."},{"key":"e_1_3_1_13_2","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR.2018.00382"},{"key":"e_1_3_1_14_2","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR.2009.5206848"},{"key":"e_1_3_1_15_2","doi-asserted-by":"publisher","DOI":"10.1109\/MSP.2012.2211477"},{"key":"e_1_3_1_16_2","first-page":"187","volume-title":"Proceedings of Pattern Recognition - 40th German Conference (GCPR)","volume":"11269","author":"D\u2019Innocente Antonio","year":"2018","unstructured":"Antonio D\u2019Innocente and Barbara Caputo. 2018. Domain Generalization with Domain-Specific Aggregation Modules. In Proceedings of Pattern Recognition - 40th German Conference (GCPR), Vol. 11269. 187\u2013198."},{"key":"e_1_3_1_17_2","first-page":"6447","volume-title":"Proceedings of Advances in Neural Information Processing Systems (NeurIPS)","author":"Dou Qi","year":"2019","unstructured":"Qi Dou, Daniel Coelho de Castro, Konstantinos Kamnitsas, and Ben Glocker. 2019. Domain Generalization via Model-Agnostic Learning of Semantic Features. In Proceedings of Advances in Neural Information Processing Systems (NeurIPS). 6447\u20136458."},{"key":"e_1_3_1_18_2","doi-asserted-by":"publisher","DOI":"10.1007\/978-3-030-58607-2_12"},{"key":"e_1_3_1_19_2","first-page":"1","volume-title":"Proceedings of 9th International Conference on Learning Representations (ICLR)","author":"Du Ying-Jun","year":"2021","unstructured":"Ying-Jun Du, Xiantong Zhen, Ling Shao, and Cees G. M. Snoek. 2021. MetaNorm: Learning to Normalize Few-Shot Batches Across Domains. In Proceedings of 9th International Conference on Learning Representations (ICLR). 1\u201313."},{"key":"e_1_3_1_20_2","doi-asserted-by":"publisher","DOI":"10.1109\/ICCV.2013.208"},{"key":"e_1_3_1_21_2","first-page":"1126","volume-title":"Proceedings of the 34th International Conference on Machine Learning (ICML)","volume":"70","author":"Finn Chelsea","year":"2017","unstructured":"Chelsea Finn, Pieter Abbeel, and Sergey Levine. 2017. Model-Agnostic Meta-Learning for Fast Adaptation of Deep Networks. In Proceedings of the 34th International Conference on Machine Learning (ICML), Vol. 70. 1126\u20131135."},{"key":"e_1_3_1_22_2","doi-asserted-by":"publisher","DOI":"10.1109\/ICCV.2015.293"},{"key":"e_1_3_1_23_2","first-page":"367","volume-title":"Proceedings of the 17th International Conference on Machine Learning (ICML)","author":"Heskes Tom","year":"2000","unstructured":"Tom Heskes. 2000. Empirical Bayes for Learning to Learn. In Proceedings of the 17th International Conference on Machine Learning (ICML). 367\u2013374."},{"issue":"9","key":"e_1_3_1_24_2","first-page":"5149","article-title":"Meta-Learning in Neural Networks: A Survey","volume":"44","author":"Hospedales Timothy M.","year":"2022","unstructured":"Timothy M. Hospedales, Antreas Antoniou, Paul Micaelli, and Amos J. Storkey. 2022. Meta-Learning in Neural Networks: A Survey. IEEE Transactions on Pattern Analysis and Machine Intelligence 44, 9 (2022), 5149\u20135169.","journal-title":"IEEE Transactions on Pattern Analysis and Machine Intelligence"},{"key":"e_1_3_1_25_2","unstructured":"Mike Huisman Jan N van Rijn and Aske Plaat. 2020. A Survey of Deep Meta-Learning. arXiv:2010.03522. Retrieved from https:\/\/arxiv.org\/abs\/2010.03522"},{"key":"e_1_3_1_26_2","first-page":"256","volume-title":"Proceedings of the 45th Annual Meeting of the Association for Computational Linguistics (ACL)","author":"III Hal Daum\u00e9","year":"2007","unstructured":"Hal Daum\u00e9 III. 2007. Frustratingly Easy Domain Adaptation. In Proceedings of the 45th Annual Meeting of the Association for Computational Linguistics (ACL). 256\u2013263."},{"key":"e_1_3_1_27_2","first-page":"1","volume-title":"Proceedings of the 3rd International Conference on Learning Representations (ICLR)","author":"Kingma Diederik P.","year":"2015","unstructured":"Diederik P. Kingma and Jimmy Ba. 2015. Adam: A Method for Stochastic Optimization. In Proceedings of the 3rd International Conference on Learning Representations (ICLR). 1\u201311."},{"key":"e_1_3_1_28_2","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR.2018.00363"},{"key":"e_1_3_1_29_2","first-page":"1","volume-title":"Proceedings of Handbook of Systemic Autoimmune Diseases","volume":"1","author":"Krizhevsky A.","year":"2009","unstructured":"A. Krizhevsky and G. Hinton. 2009. Learning Multiple Layers of Features from Tiny Images. Proceedings of Handbook of Systemic Autoimmune Diseases 1, 4 (2009). 1\u201347."},{"key":"e_1_3_1_30_2","doi-asserted-by":"publisher","DOI":"10.1007\/978-3-030-66415-2_39"},{"key":"e_1_3_1_31_2","doi-asserted-by":"publisher","DOI":"10.1109\/ICCV.2017.591"},{"key":"e_1_3_1_32_2","unstructured":"Da Li Yongxin Yang Yi-Zhe Song and Timothy M. Hospedales. 2017b. Learning to Generalize: Meta-Learning for Domain Generalization. arXiv:1710.03463. Retrieved from http:\/\/arxiv.org\/abs\/1710.03463"},{"key":"e_1_3_1_33_2","doi-asserted-by":"publisher","DOI":"10.1109\/ICCV.2019.00153"},{"key":"e_1_3_1_34_2","unstructured":"Haoyang Li Xin Wang Ziwei Zhang and Wenwu Zhu. 2022. Out-of-Distribution Generalization on Graphs: A Survey. arXiv:2202.07987. Retrieved from https:\/\/arxiv.org\/abs\/2202.07987"},{"key":"e_1_3_1_35_2","unstructured":"Yiying Li Yongxin Yang Wei Zhou and Timothy M Hospedales. 2019a. Feature-Critic Networks for Heterogeneous Domain Generalization. arXiv:1901.11448. Retrieved from http:\/\/arxiv.org\/abs\/1901.11448"},{"key":"e_1_3_1_36_2","first-page":"6781","volume-title":"Proceedings of the 38th International Conference on Machine Learning (ICML)","volume":"139","author":"Liu Evan Zheran","year":"2021","unstructured":"Evan Zheran Liu, Behzad Haghgoo, Annie S. Chen, Aditi Raghunathan, Pang Wei Koh, Shiori Sagawa, Percy Liang, and Chelsea Finn. 2021. Just Train Twice: Improving Group Robustness without Training Group Information. In Proceedings of the 38th International Conference on Machine Learning (ICML), Vol. 139. PMLR, 6781\u20136792."},{"key":"e_1_3_1_37_2","first-page":"1","volume-title":"Proceedings of 7th International Conference on Learning Representations (ICLR)","author":"Liu Yanbin","year":"2019","unstructured":"Yanbin Liu, Juho Lee, Minseop Park, Saehoon Kim, Eunho Yang, Sung Ju Hwang, and Yi Yang. 2019. Learning to Propagate Labels: Transductive Propagation Network for Few-Shot Learning. In Proceedings of 7th International Conference on Learning Representations (ICLR). 1\u201311."},{"key":"e_1_3_1_38_2","first-page":"1","volume-title":"Proceedings of 6th International Conference on Learning Representations (ICLR)","author":"Mishra Nikhil","year":"2018","unstructured":"Nikhil Mishra, Mostafa Rohaninejad, Xi Chen, and Pieter Abbeel. 2018. A Simple Neural Attentive Meta-Learner. In Proceedings of 6th International Conference on Learning Representations (ICLR). 1\u201312."},{"key":"e_1_3_1_39_2","doi-asserted-by":"publisher","DOI":"10.1109\/ICCV.2017.609"},{"key":"e_1_3_1_40_2","first-page":"2067","volume-title":"Proceedings of the Advances in Neural Information Processing Systems 33: Annual Conference on Neural Information Processing Systems (NeurIPS)","author":"Nam Jun Hyun","year":"2020","unstructured":"Jun Hyun Nam, Hyuntak Cha, Sungsoo Ahn, Jaeho Lee, and Jinwoo Shin. 2020. Learning from Failure: De-Biasing Classifier from Biased Classifier. In Proceedings of the Advances in Neural Information Processing Systems 33: Annual Conference on Neural Information Processing Systems (NeurIPS). 2067\u201320684."},{"key":"e_1_3_1_41_2","first-page":"2850","volume-title":"Proceedings of Advances in Neural Information Processing Systems (NeurIPS)","author":"Schwartz Eli","year":"2018","unstructured":"Eli Schwartz, Leonid Karlinsky, Joseph Shtok, Sivan Harary, Mattias Marder, Abhishek Kumar, Rog\u00e9rio Schmidt Feris, Raja Giryes, and Alexander M. Bronstein. 2018. Delta-Encoder: An Effective Sample Synthesis Method for Few-Shot Object Recognition. In Proceedings of Advances in Neural Information Processing Systems (NeurIPS). 2850\u20132860."},{"key":"e_1_3_1_42_2","unstructured":"Jian Shen Yanru Qu Weinan Zhang and Yong Yu. 2017. Wasserstein Distance Guided Representation Learning for Domain Adaptation. arXiv:1707.01217. Retrieved from https:\/\/arxiv.org\/abs\/1707.01217"},{"key":"e_1_3_1_43_2","unstructured":"Zheyan Shen Jiashuo Liu Yue He Xingxuan Zhang Renzhe Xu Han Yu and Peng Cui. 2021. Towards Out-of-Distribution Generalization: A Survey. arXiv:2108.13624. Retrieved from https:\/\/arxiv.org\/abs\/2108.13624"},{"key":"e_1_3_1_44_2","doi-asserted-by":"publisher","DOI":"10.1145\/3534678.3539249"},{"key":"e_1_3_1_45_2","doi-asserted-by":"publisher","DOI":"10.1109\/ICCV.2017.386"},{"key":"e_1_3_1_46_2","first-page":"4790","volume-title":"Proceedings of Advances in Neural Information Processing Systems (NeurIPS)","author":"Oord A\u00e4ron van den","year":"2016","unstructured":"A\u00e4ron van den Oord, Nal Kalchbrenner, Lasse Espeholt, Koray Kavukcuoglu, Oriol Vinyals, and Alex Graves. 2016. Conditional Image Generation with PixelCNN Decoders. In Proceedings of Advances in Neural Information Processing Systems (NeurIPS). 4790\u20134798."},{"key":"e_1_3_1_47_2","first-page":"6904","volume-title":"Proceedings of Advances in Neural Information Processing Systems (NeurIPS)","author":"Vartak Manasi","year":"2017","unstructured":"Manasi Vartak, Arvind Thiagarajan, Conrado Miranda, Jeshua Bratman, and Hugo Larochelle. 2017. A Meta-Learning Perspective on Cold-Start Recommendations for Items. In Proceedings of Advances in Neural Information Processing Systems (NeurIPS). 6904\u20136914."},{"key":"e_1_3_1_48_2","doi-asserted-by":"publisher","DOI":"10.24963\/ijcai.2021\/628"},{"key":"e_1_3_1_49_2","doi-asserted-by":"publisher","DOI":"10.24963\/ijcai.2022\/508"},{"key":"e_1_3_1_50_2","first-page":"26484","volume-title":"Proceedings of the International Conference on Machine Learning (ICML)","volume":"162","author":"Zhang Michael","year":"2022","unstructured":"Michael Zhang, Nimit Sharad Sohoni, Hongyang R. Zhang, Chelsea Finn, and Christopher R\u00e9. 2022. Correct-N-Contrast: A Contrastive Approach for Improving Robustness to Spurious Correlations. In Proceedings of the International Conference on Machine Learning (ICML), Vol. 162. PMLR, 26484\u201326516."},{"key":"e_1_3_1_51_2","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR46437.2021.00533"},{"key":"e_1_3_1_52_2","unstructured":"Kaiyang Zhou Ziwei Liu Yu Qiao Tao Xiang and Chen Change Loy. 2021. Domain Generalization: A Survey. arXiv:2103.02503. Retrieved from https:\/\/arxiv.org\/abs\/2103.02503"},{"key":"e_1_3_1_53_2","first-page":"27222","volume-title":"Proceedings of the International Conference on Machine Learning (ICML)","volume":"162","author":"Zhou Xiao","year":"2022","unstructured":"Xiao Zhou, Yong Lin, Weizhong Zhang, and Tong Zhang. 2022. Sparse Invariant Risk Minimization. In Proceedings of the International Conference on Machine Learning (ICML), Vol. 162. 27222\u201327244."}],"container-title":["ACM Transactions on Knowledge Discovery from Data"],"original-title":[],"language":"en","link":[{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/3676558","content-type":"unspecified","content-version":"vor","intended-application":"text-mining"},{"URL":"https:\/\/dl.acm.org\/doi\/pdf\/10.1145\/3676558","content-type":"unspecified","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2025,6,19]],"date-time":"2025-06-19T01:18:46Z","timestamp":1750295926000},"score":1,"resource":{"primary":{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/3676558"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2024,8,16]]},"references-count":52,"journal-issue":{"issue":"8","published-print":{"date-parts":[[2024,9,30]]}},"alternative-id":["10.1145\/3676558"],"URL":"https:\/\/doi.org\/10.1145\/3676558","relation":{},"ISSN":["1556-4681","1556-472X"],"issn-type":[{"value":"1556-4681","type":"print"},{"value":"1556-472X","type":"electronic"}],"subject":[],"published":{"date-parts":[[2024,8,16]]},"assertion":[{"value":"2022-11-25","order":0,"name":"received","label":"Received","group":{"name":"publication_history","label":"Publication History"}},{"value":"2024-06-27","order":2,"name":"accepted","label":"Accepted","group":{"name":"publication_history","label":"Publication History"}},{"value":"2024-08-16","order":3,"name":"published","label":"Published","group":{"name":"publication_history","label":"Publication History"}}]}}