{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2025,6,18]],"date-time":"2025-06-18T04:09:12Z","timestamp":1750219752760,"version":"3.41.0"},"publisher-location":"New York, NY, USA","reference-count":38,"publisher":"ACM","license":[{"start":{"date-parts":[[2023,10,21]],"date-time":"2023-10-21T00:00:00Z","timestamp":1697846400000},"content-version":"vor","delay-in-days":0,"URL":"https:\/\/www.acm.org\/publications\/policies\/copyright_policy#Background"}],"content-domain":{"domain":["dl.acm.org"],"crossmark-restriction":true},"short-container-title":[],"published-print":{"date-parts":[[2023,10,21]]},"DOI":"10.1145\/3583780.3615462","type":"proceedings-article","created":{"date-parts":[[2023,10,21]],"date-time":"2023-10-21T07:45:42Z","timestamp":1697874342000},"page":"4559-4566","update-policy":"https:\/\/doi.org\/10.1145\/crossmark-policy","source":"Crossref","is-referenced-by-count":0,"title":["Content-Based Email Classification at Scale"],"prefix":"10.1145","author":[{"ORCID":"https:\/\/orcid.org\/0000-0002-1627-0094","authenticated-orcid":false,"given":"Kirstin","family":"Early","sequence":"first","affiliation":[{"name":"Yahoo Research, Mountain View, CA, USA"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"ORCID":"https:\/\/orcid.org\/0009-0001-7499-7814","authenticated-orcid":false,"given":"Neil","family":"O'Hare","sequence":"additional","affiliation":[{"name":"Yahoo Research, San Francisco, CA, USA"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"ORCID":"https:\/\/orcid.org\/0009-0002-2567-1305","authenticated-orcid":false,"given":"Christopher","family":"Luvogt","sequence":"additional","affiliation":[{"name":"Yahoo Research, Mountain View, CA, USA"}],"role":[{"role":"author","vocabulary":"crossref"}]}],"member":"320","published-online":{"date-parts":[[2023,10,21]]},"reference":[{"volume-title":"Large Scale Distributed Neural Network Training Through Online Distillation. In International Conference on Learning Representations. https:\/\/openreview.net\/forum?id=rkr1UDeC-","author":"Anil Rohan","unstructured":"Rohan Anil , Gabriel Pereyra , Alexandre Passos , Robert Ormandi , George E. Dahl , and Geoffrey E. Hinton . 2018 . Large Scale Distributed Neural Network Training Through Online Distillation. In International Conference on Learning Representations. https:\/\/openreview.net\/forum?id=rkr1UDeC- Rohan Anil, Gabriel Pereyra, Alexandre Passos, Robert Ormandi, George E. Dahl, and Geoffrey E. Hinton. 2018. Large Scale Distributed Neural Network Training Through Online Distillation. In International Conference on Learning Representations. https:\/\/openreview.net\/forum?id=rkr1UDeC-","key":"e_1_3_2_1_1_1"},{"doi-asserted-by":"publisher","key":"e_1_3_2_1_2_1","DOI":"10.18653\/v1\/P19-1633"},{"doi-asserted-by":"publisher","key":"e_1_3_2_1_4_1","DOI":"10.1162\/tacl_a_00051"},{"key":"e_1_3_2_1_5_1","volume-title":"Model Compression. In Proceedings of the 12th ACM SIGKDD International Conference on Knowledge Discovery and Data Mining (KDD '06)","author":"Cristian","year":"2006","unstructured":"Cristian Bucilu?, Rich Caruana , and Alexandru Niculescu-Mizil . 2006 . Model Compression. In Proceedings of the 12th ACM SIGKDD International Conference on Knowledge Discovery and Data Mining (KDD '06) . Association for Computing Machinery, New York, NY, USA, 535--541. https:\/\/doi.org\/10.1145\/1150402.1150464 10.1145\/1150402.1150464 Cristian Bucilu?, Rich Caruana, and Alexandru Niculescu-Mizil. 2006. Model Compression. In Proceedings of the 12th ACM SIGKDD International Conference on Knowledge Discovery and Data Mining (KDD '06). Association for Computing Machinery, New York, NY, USA, 535--541. https:\/\/doi.org\/10.1145\/1150402.1150464"},{"unstructured":"California State Legislature. 2018. California Consumer Privacy Act. newlinehttps:\/\/oag.ca.gov\/privacy\/ccpa.  California State Legislature. 2018. California Consumer Privacy Act. newlinehttps:\/\/oag.ca.gov\/privacy\/ccpa.","key":"e_1_3_2_1_6_1"},{"volume-title":"International Conference on Learning Representations. https:\/\/openreview.net\/forum?id=r1xMH1BtvB","author":"Clark Kevin","unstructured":"Kevin Clark , Minh-Thang Luong , Quoc V. Le , and Christopher D. Manning . 2020. ELECTRA: Pre-training Text Encoders as Discriminators Rather Than Generators . In International Conference on Learning Representations. https:\/\/openreview.net\/forum?id=r1xMH1BtvB Kevin Clark, Minh-Thang Luong, Quoc V. Le, and Christopher D. Manning. 2020. ELECTRA: Pre-training Text Encoders as Discriminators Rather Than Generators. In International Conference on Learning Representations. https:\/\/openreview.net\/forum?id=r1xMH1BtvB","key":"e_1_3_2_1_7_1"},{"key":"e_1_3_2_1_8_1","volume-title":"Total Audience","author":"Metrix\u00ae Multi-Platform Comscore Media","year":"2023","unstructured":"Comscore Media Metrix\u00ae Multi-Platform . 2023 . Services - e-mail category , Total Audience , June 2023, U.S. Comscore Media Metrix\u00ae Multi-Platform. 2023. Services - e-mail category, Total Audience, June 2023, U.S."},{"unstructured":"Council of European Union. 2016. Regulation (EU) 2016\/679 of the European Parliament and of the Council. newlinehttp:\/\/data.europa.eu\/eli\/reg\/2016\/679\/oj.  Council of European Union. 2016. Regulation (EU) 2016\/679 of the European Parliament and of the Council. newlinehttp:\/\/data.europa.eu\/eli\/reg\/2016\/679\/oj.","key":"e_1_3_2_1_9_1"},{"key":"e_1_3_2_1_10_1","volume-title":"Proceedings of the 2019 Conference of the North American Chapter of the Association for Computational Linguistics: Human Language Technologies","volume":"1","author":"Devlin Jacob","year":"2019","unstructured":"Jacob Devlin , Ming-Wei Chang , Kenton Lee , and Kristina Toutanova . 2019 . BERT: Pre-training of Deep Bidirectional Transformers for Language Understanding . In Proceedings of the 2019 Conference of the North American Chapter of the Association for Computational Linguistics: Human Language Technologies , Volume 1 (Long and Short Papers). Association for Computational Linguistics, Minneapolis, Minnesota, 4171--4186. https:\/\/doi.org\/10. 18653\/v1\/N19--1423 10.18653\/v1 Jacob Devlin, Ming-Wei Chang, Kenton Lee, and Kristina Toutanova. 2019. BERT: Pre-training of Deep Bidirectional Transformers for Language Understanding. In Proceedings of the 2019 Conference of the North American Chapter of the Association for Computational Linguistics: Human Language Technologies, Volume 1 (Long and Short Papers). Association for Computational Linguistics, Minneapolis, Minnesota, 4171--4186. https:\/\/doi.org\/10.18653\/v1\/N19--1423"},{"doi-asserted-by":"publisher","key":"e_1_3_2_1_11_1","DOI":"10.1007\/s11263-021-01453-z"},{"doi-asserted-by":"publisher","key":"e_1_3_2_1_12_1","DOI":"10.1145\/2661829.2662018"},{"key":"e_1_3_2_1_13_1","volume-title":"NIPS Deep Learning and Representation Learning Workshop. http:\/\/arxiv.org\/abs\/1503","author":"Hinton Geoffrey","year":"2015","unstructured":"Geoffrey Hinton , Oriol Vinyals , and Jeffrey Dean . 2015 . Distilling the Knowledge in a Neural Network . In NIPS Deep Learning and Representation Learning Workshop. http:\/\/arxiv.org\/abs\/1503 .02531 Geoffrey Hinton, Oriol Vinyals, and Jeffrey Dean. 2015. Distilling the Knowledge in a Neural Network. In NIPS Deep Learning and Representation Learning Workshop. http:\/\/arxiv.org\/abs\/1503.02531"},{"key":"e_1_3_2_1_14_1","first-page":"372","volume-title":"TinyBERT: Distilling BERT for Natural Language Understanding. In Findings of the Association for Computational Linguistics: EMNLP 2020","author":"Jiao Xiaoqi","year":"2020","unstructured":"Xiaoqi Jiao , Yichun Yin , Lifeng Shang , Xin Jiang , Xiao Chen , Linlin Li , Fang Wang , and Qun Liu . 2020 . TinyBERT: Distilling BERT for Natural Language Understanding. In Findings of the Association for Computational Linguistics: EMNLP 2020 . Association for Computational Linguistics, Online, 4163--4174. https:\/\/doi.org\/10. 18653\/v1\/2020.findings-emnlp. 372 10.18653\/v1 Xiaoqi Jiao, Yichun Yin, Lifeng Shang, Xin Jiang, Xiao Chen, Linlin Li, Fang Wang, and Qun Liu. 2020. TinyBERT: Distilling BERT for Natural Language Understanding. In Findings of the Association for Computational Linguistics: EMNLP 2020. Association for Computational Linguistics, Online, 4163--4174. https:\/\/doi.org\/10.18653\/v1\/2020.findings-emnlp.372"},{"doi-asserted-by":"publisher","key":"e_1_3_2_1_15_1","DOI":"10.18653\/v1\/E17-2068"},{"doi-asserted-by":"publisher","key":"e_1_3_2_1_16_1","DOI":"10.1609\/aaai.v36i7.20666"},{"doi-asserted-by":"publisher","key":"e_1_3_2_1_17_1","DOI":"10.1145\/2020408.2020560"},{"key":"e_1_3_2_1_18_1","volume-title":"Proceedings of the 2018 Conference on Empirical Methods in Natural Language Processing: System Demonstrations. Association for Computational Linguistics","author":"Kudo Taku","year":"2018","unstructured":"Taku Kudo and John Richardson . 2018 . SentencePiece: A Simple and Language Independent Subword Tokenizer and Detokenizer for Neural Text Processing . In Proceedings of the 2018 Conference on Empirical Methods in Natural Language Processing: System Demonstrations. Association for Computational Linguistics , Brussels, Belgium, 66--71. https:\/\/doi.org\/10. 18653\/v1\/D18--2012 10.18653\/v1 Taku Kudo and John Richardson. 2018. SentencePiece: A Simple and Language Independent Subword Tokenizer and Detokenizer for Neural Text Processing. In Proceedings of the 2018 Conference on Empirical Methods in Natural Language Processing: System Demonstrations. Association for Computational Linguistics, Brussels, Belgium, 66--71. https:\/\/doi.org\/10.18653\/v1\/D18--2012"},{"key":"e_1_3_2_1_19_1","volume-title":"ALBERT: A Lite BERT for Self-supervised Learning of Language Representations. In 8th International Conference on Learning Representations, ICLR 2020","author":"Lan Zhenzhong","year":"2020","unstructured":"Zhenzhong Lan , Mingda Chen , Sebastian Goodman , Kevin Gimpel , Piyush Sharma , and Radu Soricut . 2020 . ALBERT: A Lite BERT for Self-supervised Learning of Language Representations. In 8th International Conference on Learning Representations, ICLR 2020 , Addis Ababa, Ethiopia, April 26--30 , 2020. OpenReview.net. https:\/\/openreview.net\/forum?id=H1eA7AEtvS Zhenzhong Lan, Mingda Chen, Sebastian Goodman, Kevin Gimpel, Piyush Sharma, and Radu Soricut. 2020. ALBERT: A Lite BERT for Self-supervised Learning of Language Representations. In 8th International Conference on Learning Representations, ICLR 2020, Addis Ababa, Ethiopia, April 26--30, 2020. OpenReview.net. https:\/\/openreview.net\/forum?id=H1eA7AEtvS"},{"key":"e_1_3_2_1_20_1","volume-title":"RoBERTa: A Robustly Optimized BERT Pretraining Approach. CoRR","author":"Liu Yinhan","year":"2019","unstructured":"Yinhan Liu , Myle Ott , Naman Goyal , Jingfei Du , Mandar Joshi , Danqi Chen , Omer Levy , Mike Lewis , Luke Zettlemoyer , and Veselin Stoyanov . 2019. RoBERTa: A Robustly Optimized BERT Pretraining Approach. CoRR , Vol. abs\/ 1907 .11692 ( 2019 ). showeprint[arXiv]1907.11692 http:\/\/arxiv.org\/abs\/1907.11692 Yinhan Liu, Myle Ott, Naman Goyal, Jingfei Du, Mandar Joshi, Danqi Chen, Omer Levy, Mike Lewis, Luke Zettlemoyer, and Veselin Stoyanov. 2019. RoBERTa: A Robustly Optimized BERT Pretraining Approach. CoRR, Vol. abs\/1907.11692 (2019). showeprint[arXiv]1907.11692 http:\/\/arxiv.org\/abs\/1907.11692"},{"key":"e_1_3_2_1_21_1","volume-title":"Conditional Teacher-Student Learning. 2019 IEEE International Conference on Acoustics, Speech and Signal Processing (ICASSP) (2019","author":"Meng Zhong","year":"2019","unstructured":"Zhong Meng , Jinyu Li , Yong Zhao , and Yifan Gong . 2019 . Conditional Teacher-Student Learning. 2019 IEEE International Conference on Acoustics, Speech and Signal Processing (ICASSP) (2019 ), 6445--6449. http:\/\/arxiv.org\/abs\/1904.12399 Zhong Meng, Jinyu Li, Yong Zhao, and Yifan Gong. 2019. Conditional Teacher-Student Learning. 2019 IEEE International Conference on Acoustics, Speech and Signal Processing (ICASSP) (2019), 6445--6449. http:\/\/arxiv.org\/abs\/1904.12399"},{"doi-asserted-by":"publisher","key":"e_1_3_2_1_22_1","DOI":"10.1609\/aaai.v35i15.17610"},{"doi-asserted-by":"publisher","key":"e_1_3_2_1_23_1","DOI":"10.1109\/TKDE.2019.2959991"},{"key":"e_1_3_2_1_24_1","volume-title":"Antoine Chassang, Carlo Gatta, and Yoshua Bengio.","author":"Romero Adriana","year":"2015","unstructured":"Adriana Romero , Nicolas Ballas , Samira Ebrahimi Kahou , Antoine Chassang, Carlo Gatta, and Yoshua Bengio. 2015 . FitNets: Hints for Thin Deep Nets. In 3rd International Conference on Learning Representations, ICLR 2015, San Diego, CA, USA, May 7--9, 2015, Conference Track Proceedings, Yoshua Bengio and Yann LeCun (Eds .). http:\/\/arxiv.org\/abs\/1412.6550 Adriana Romero, Nicolas Ballas, Samira Ebrahimi Kahou, Antoine Chassang, Carlo Gatta, and Yoshua Bengio. 2015. FitNets: Hints for Thin Deep Nets. In 3rd International Conference on Learning Representations, ICLR 2015, San Diego, CA, USA, May 7--9, 2015, Conference Track Proceedings, Yoshua Bengio and Yann LeCun (Eds.). http:\/\/arxiv.org\/abs\/1412.6550"},{"key":"e_1_3_2_1_25_1","volume-title":"Faster, Cheaper and Lighter. CoRR","author":"Sanh Victor","year":"2019","unstructured":"Victor Sanh , Lysandre Debut , Julien Chaumond , and Thomas Wolf . 2019. DistilBERT , a Distilled Version of BERT: Smaller , Faster, Cheaper and Lighter. CoRR , Vol. abs\/ 1910 .01108 ( 2019 ). showeprint[arXiv]1910.01108 http:\/\/arxiv.org\/abs\/1910.01108 Victor Sanh, Lysandre Debut, Julien Chaumond, and Thomas Wolf. 2019. DistilBERT, a Distilled Version of BERT: Smaller, Faster, Cheaper and Lighter. CoRR, Vol. abs\/1910.01108 (2019). showeprint[arXiv]1910.01108 http:\/\/arxiv.org\/abs\/1910.01108"},{"doi-asserted-by":"publisher","key":"e_1_3_2_1_26_1","DOI":"10.18653\/v1\/P16-1162"},{"volume-title":"Number of sent and received e-mails per day worldwide from 2017 to","year":"2025","unstructured":"Statista. 2022. Number of sent and received e-mails per day worldwide from 2017 to 2025 . https:\/\/www.statista.com\/statistics\/456500\/daily-number-of-e-mails-worldwide\/ Statista. 2022. Number of sent and received e-mails per day worldwide from 2017 to 2025. https:\/\/www.statista.com\/statistics\/456500\/daily-number-of-e-mails-worldwide\/","key":"e_1_3_2_1_27_1"},{"doi-asserted-by":"publisher","key":"e_1_3_2_1_28_1","DOI":"10.18653\/v1\/2020.acl-main.195"},{"doi-asserted-by":"publisher","key":"e_1_3_2_1_29_1","DOI":"10.4018\/978-1-60566-058-5.ch021"},{"key":"e_1_3_2_1_30_1","volume-title":"Well-Read Students Learn Better: The Impact of Student Initialization on Knowledge Distillation. CoRR","author":"Turc Iulia","year":"2019","unstructured":"Iulia Turc , Ming-Wei Chang , Kenton Lee , and Kristina Toutanova . 2019. Well-Read Students Learn Better: The Impact of Student Initialization on Knowledge Distillation. CoRR , Vol. abs\/ 1908 .08962 ( 2019 ). showeprint[arXiv]1908.08962 http:\/\/arxiv.org\/abs\/1908.08962 Iulia Turc, Ming-Wei Chang, Kenton Lee, and Kristina Toutanova. 2019. Well-Read Students Learn Better: The Impact of Student Initialization on Knowledge Distillation. CoRR, Vol. abs\/1908.08962 (2019). showeprint[arXiv]1908.08962 http:\/\/arxiv.org\/abs\/1908.08962"},{"key":"e_1_3_2_1_31_1","first-page":"I","article-title":"Attention Is All You Need","volume":"30","author":"Vaswani Ashish","year":"2017","unstructured":"Ashish Vaswani , Noam Shazeer , Niki Parmar , Jakob Uszkoreit , Llion Jones , Aidan N. Gomez , \u0141ukasz Kaiser , and Illia Polosukhin . 2017 . Attention Is All You Need . In Advances in Neural Information Processing Systems 30 , I . Guyon, U. V. Luxburg, S. Bengio, H. Wallach, R. Fergus, S. Vishwanathan, and R. Garnett (Eds.). Curran Associates, Inc., 5998--6008. http:\/\/papers.nips.cc\/paper\/7181-attention-is-all-you-need.pdf Ashish Vaswani, Noam Shazeer, Niki Parmar, Jakob Uszkoreit, Llion Jones, Aidan N. Gomez, \u0141ukasz Kaiser, and Illia Polosukhin. 2017. Attention Is All You Need. In Advances in Neural Information Processing Systems 30, I. Guyon, U. V. Luxburg, S. Bengio, H. Wallach, R. Fergus, S. Vishwanathan, and R. Garnett (Eds.). Curran Associates, Inc., 5998--6008. http:\/\/papers.nips.cc\/paper\/7181-attention-is-all-you-need.pdf","journal-title":"Advances in Neural Information Processing Systems"},{"key":"e_1_3_2_1_32_1","volume-title":"Advances in Neural Information Processing Systems","author":"Wang Wenhui","year":"2020","unstructured":"Wenhui Wang , Furu Wei , Li Dong , Hangbo Bao , Nan Yang , and Ming Zhou . 2020. MiniLM: Deep Self-Attention Distillation for Task-Agnostic Compression of Pre-Trained Transformers . In Advances in Neural Information Processing Systems , H. Larochelle, M. Ranzato, R. Hadsell, M.F. Balcan, and H. Lin (Eds.), Vol. 33 . Curran Associates, Inc. , 5776--5788. https:\/\/proceedings.neurips.cc\/paper_files\/paper\/ 2020 \/file\/3f5ee243547dee91fbd053c1c4a845aa-Paper.pdf Wenhui Wang, Furu Wei, Li Dong, Hangbo Bao, Nan Yang, and Ming Zhou. 2020. MiniLM: Deep Self-Attention Distillation for Task-Agnostic Compression of Pre-Trained Transformers. In Advances in Neural Information Processing Systems, H. Larochelle, M. Ranzato, R. Hadsell, M.F. Balcan, and H. Lin (Eds.), Vol. 33. Curran Associates, Inc., 5776--5788. https:\/\/proceedings.neurips.cc\/paper_files\/paper\/2020\/file\/3f5ee243547dee91fbd053c1c4a845aa-Paper.pdf"},{"doi-asserted-by":"publisher","key":"e_1_3_2_1_33_1","DOI":"10.1145\/2835776.2835780"},{"key":"e_1_3_2_1_34_1","volume-title":"Google's Neural Machine Translation System: Bridging the Gap between Human and Machine Translation. CoRR","author":"Wu Yonghui","year":"2016","unstructured":"Yonghui Wu , Mike Schuster , Zhifeng Chen , Quoc V. Le , Mohammad Norouzi , Wolfgang Macherey , Maxim Krikun , Yuan Cao , Qin Gao , Klaus Macherey , Jeff Klingner , Apurva Shah , Melvin Johnson , Xiaobing Liu , ?ukasz Kaiser, Stephan Gouws , Yoshikiyo Kato , Taku Kudo , Hideto Kazawa , Keith Stevens , George Kurian , Nishant Patil , Wei Wang , Cliff Young , Jason Smith , Jason Riesa , Alex Rudnick , Oriol Vinyals , Greg Corrado , Macduff Hughes , and Jeffrey Dean . 2016. Google's Neural Machine Translation System: Bridging the Gap between Human and Machine Translation. CoRR , Vol. abs\/ 1609 .08144 ( 2016 ). http:\/\/arxiv.org\/abs\/1609.08144 Yonghui Wu, Mike Schuster, Zhifeng Chen, Quoc V. Le, Mohammad Norouzi, Wolfgang Macherey, Maxim Krikun, Yuan Cao, Qin Gao, Klaus Macherey, Jeff Klingner, Apurva Shah, Melvin Johnson, Xiaobing Liu, ?ukasz Kaiser, Stephan Gouws, Yoshikiyo Kato, Taku Kudo, Hideto Kazawa, Keith Stevens, George Kurian, Nishant Patil, Wei Wang, Cliff Young, Jason Smith, Jason Riesa, Alex Rudnick, Oriol Vinyals, Greg Corrado, Macduff Hughes, and Jeffrey Dean. 2016. Google's Neural Machine Translation System: Bridging the Gap between Human and Machine Translation. CoRR, Vol. abs\/1609.08144 (2016). http:\/\/arxiv.org\/abs\/1609.08144"},{"doi-asserted-by":"publisher","key":"e_1_3_2_1_35_1","DOI":"10.1145\/3534678.3539189"},{"key":"e_1_3_2_1_36_1","volume-title":"Network Minimization and Transfer Learning. In 2017 IEEE Conference on Computer Vision and Pattern Recognition (CVPR). 7130--7138","author":"Yim Junho","year":"2017","unstructured":"Junho Yim , Donggyu Joo , Jihoon Bae , and Junmo Kim . 2017 . A Gift from Knowledge Distillation: Fast Optimization , Network Minimization and Transfer Learning. In 2017 IEEE Conference on Computer Vision and Pattern Recognition (CVPR). 7130--7138 . https:\/\/doi.org\/10.1109\/CVPR.2017.754 10.1109\/CVPR.2017.754 Junho Yim, Donggyu Joo, Jihoon Bae, and Junmo Kim. 2017. A Gift from Knowledge Distillation: Fast Optimization, Network Minimization and Transfer Learning. In 2017 IEEE Conference on Computer Vision and Pattern Recognition (CVPR). 7130--7138. https:\/\/doi.org\/10.1109\/CVPR.2017.754"},{"doi-asserted-by":"publisher","key":"e_1_3_2_1_37_1","DOI":"10.24963\/ijcai.2018\/158"},{"key":"e_1_3_2_1_38_1","volume-title":"Deep Mutual Learning. In 2018 IEEE\/CVF Conference on Computer Vision and Pattern Recognition. 4320--4328","author":"Zhang Ying","year":"2018","unstructured":"Ying Zhang , Tao Xiang , Timothy M. Hospedales , and Huchuan Lu . 2018 . Deep Mutual Learning. In 2018 IEEE\/CVF Conference on Computer Vision and Pattern Recognition. 4320--4328 . https:\/\/doi.org\/10.1109\/CVPR.2018.00454 10.1109\/CVPR.2018.00454 Ying Zhang, Tao Xiang, Timothy M. Hospedales, and Huchuan Lu. 2018. Deep Mutual Learning. In 2018 IEEE\/CVF Conference on Computer Vision and Pattern Recognition. 4320--4328. https:\/\/doi.org\/10.1109\/CVPR.2018.00454"},{"doi-asserted-by":"publisher","key":"e_1_3_2_1_39_1","DOI":"10.18653\/v1\/2020.acl-main.104"}],"event":{"sponsor":["SIGWEB ACM Special Interest Group on Hypertext, Hypermedia, and Web","SIGIR ACM Special Interest Group on Information Retrieval"],"acronym":"CIKM '23","name":"CIKM '23: The 32nd ACM International Conference on Information and Knowledge Management","location":"Birmingham United Kingdom"},"container-title":["Proceedings of the 32nd ACM International Conference on Information and Knowledge Management"],"original-title":[],"link":[{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/3583780.3615462","content-type":"unspecified","content-version":"vor","intended-application":"text-mining"},{"URL":"https:\/\/dl.acm.org\/doi\/pdf\/10.1145\/3583780.3615462","content-type":"unspecified","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2025,6,17]],"date-time":"2025-06-17T16:36:54Z","timestamp":1750178214000},"score":1,"resource":{"primary":{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/3583780.3615462"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2023,10,21]]},"references-count":38,"alternative-id":["10.1145\/3583780.3615462","10.1145\/3583780"],"URL":"https:\/\/doi.org\/10.1145\/3583780.3615462","relation":{},"subject":[],"published":{"date-parts":[[2023,10,21]]},"assertion":[{"value":"2023-10-21","order":2,"name":"published","label":"Published","group":{"name":"publication_history","label":"Publication History"}}]}}