{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2026,5,2]],"date-time":"2026-05-02T15:22:56Z","timestamp":1777735376569,"version":"3.51.4"},"publisher-location":"New York, NY, USA","reference-count":47,"publisher":"ACM","license":[{"start":{"date-parts":[[2021,8,14]],"date-time":"2021-08-14T00:00:00Z","timestamp":1628899200000},"content-version":"vor","delay-in-days":0,"URL":"https:\/\/creativecommons.org\/licenses\/by\/4.0\/"}],"funder":[{"DOI":"10.13039\/100000001","name":"NSF (National Science Foundation)","doi-asserted-by":"publisher","award":["IIS-1747614"],"award-info":[{"award-number":["IIS-1747614"]}],"id":[{"id":"10.13039\/100000001","id-type":"DOI","asserted-by":"publisher"}]}],"content-domain":{"domain":["dl.acm.org"],"crossmark-restriction":true},"short-container-title":[],"published-print":{"date-parts":[[2021,8,14]]},"DOI":"10.1145\/3447548.3467235","type":"proceedings-article","created":{"date-parts":[[2021,8,13]],"date-time":"2021-08-13T18:33:08Z","timestamp":1628879588000},"page":"1737-1747","update-policy":"https:\/\/doi.org\/10.1145\/crossmark-policy","source":"Crossref","is-referenced-by-count":27,"title":["Meta Self-training for Few-shot Neural Sequence Labeling"],"prefix":"10.1145","author":[{"given":"Yaqing","family":"Wang","sequence":"first","affiliation":[{"name":"Purdue University, West Lafayette, IN, USA"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Subhabrata","family":"Mukherjee","sequence":"additional","affiliation":[{"name":"Microsoft Research, Redmond, WA, USA"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Haoda","family":"Chu","sequence":"additional","affiliation":[{"name":"Microsoft AI, Bellevue, WA, USA"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Yuancheng","family":"Tu","sequence":"additional","affiliation":[{"name":"Microsoft AI, Bellevue, WA, USA"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Ming","family":"Wu","sequence":"additional","affiliation":[{"name":"Microsoft AI, Bellevue, WA, USA"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Jing","family":"Gao","sequence":"additional","affiliation":[{"name":"Purdue University, West Lafayette, IN, USA"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Ahmed Hassan","family":"Awadallah","sequence":"additional","affiliation":[{"name":"Microsoft Research, Redmond, WA, USA"}],"role":[{"role":"author","vocabulary":"crossref"}]}],"member":"320","published-online":{"date-parts":[[2021,8,14]]},"reference":[{"key":"e_1_3_2_1_1_1","doi-asserted-by":"publisher","DOI":"10.5555\/3157382.3157543"},{"key":"e_1_3_2_1_2_1","doi-asserted-by":"publisher","DOI":"10.1145\/1553374.1553380"},{"key":"e_1_3_2_1_3_1","first-page":"2017","article-title":"Active Bias: Training More Accurate Neural Networks by Emphasizing High Variance Samples","volume":"30","author":"Chang Haw-Shiuan","year":"2017","unstructured":"Haw-Shiuan Chang , Erik G. Learned-Miller , and Andrew McCallum . 2017 . Active Bias: Training More Accurate Neural Networks by Emphasizing High Variance Samples . In Advances in Neural Information Processing Systems 30 , 2017 . Haw-Shiuan Chang, Erik G. Learned-Miller, and Andrew McCallum. 2017. Active Bias: Training More Accurate Neural Networks by Emphasizing High Variance Samples. In Advances in Neural Information Processing Systems 30, 2017.","journal-title":"Advances in Neural Information Processing Systems"},{"key":"e_1_3_2_1_4_1","unstructured":"Olivier Chapelle Bernhard Schlkopf and Alexander Zien. 2010. Semi-Supervised Learning. (2010).  Olivier Chapelle Bernhard Schlkopf and Alexander Zien. 2010. Semi-Supervised Learning. (2010)."},{"key":"e_1_3_2_1_5_1","doi-asserted-by":"publisher","DOI":"10.18653\/v1\/2020.acl-main.777"},{"key":"e_1_3_2_1_6_1","volume-title":"Variational sequential labelers for semi-supervised learning. arXiv preprint arXiv:1906.09535","author":"Chen Mingda","year":"2019","unstructured":"Mingda Chen , Qingming Tang , Karen Livescu , and Kevin Gimpel . 2019. Variational sequential labelers for semi-supervised learning. arXiv preprint arXiv:1906.09535 ( 2019 ). Mingda Chen, Qingming Tang, Karen Livescu, and Kevin Gimpel. 2019. Variational sequential labelers for semi-supervised learning. arXiv preprint arXiv:1906.09535 (2019)."},{"key":"e_1_3_2_1_7_1","volume-title":"Semi-supervised sequence modeling with cross-view training. arXiv preprint arXiv:1809.08370","author":"Clark Kevin","year":"2018","unstructured":"Kevin Clark , Minh-Thang Luong , Christopher D Manning , and Quoc V Le. 2018. Semi-supervised sequence modeling with cross-view training. arXiv preprint arXiv:1809.08370 ( 2018 ). Kevin Clark, Minh-Thang Luong, Christopher D Manning, and Quoc V Le. 2018. Semi-supervised sequence modeling with cross-view training. arXiv preprint arXiv:1809.08370 (2018)."},{"key":"e_1_3_2_1_8_1","volume-title":"Privacy in Machine Learning and Artificial Intelligence workshop, ICML2018","author":"Coucke Alice","year":"2018","unstructured":"Alice Coucke , Alaa Saade , Adrien Ball , Th\u00e9 odore Bluche , Alexandre Caulier , David Leroy , Cl\u00e9 ment Doumouro , Thibault Gisselbrecht , Francesco Caltagirone , Thibaut Lavril , Ma\u00ebl Primet , and Joseph Dureau . 2018 . Snips Voice Platform: an embedded Spoken Language Understanding system for private-by-design voice interfaces . In Privacy in Machine Learning and Artificial Intelligence workshop, ICML2018 . Alice Coucke, Alaa Saade, Adrien Ball, Th\u00e9 odore Bluche, Alexandre Caulier, David Leroy, Cl\u00e9 ment Doumouro, Thibault Gisselbrecht, Francesco Caltagirone, Thibaut Lavril, Ma\u00ebl Primet, and Joseph Dureau. 2018. Snips Voice Platform: an embedded Spoken Language Understanding system for private-by-design voice interfaces. In Privacy in Machine Learning and Artificial Intelligence workshop, ICML2018."},{"key":"e_1_3_2_1_9_1","volume-title":"Proceedings of the 2019 Conference of the North American Chapter of the Association for Computational Linguistics: Human Language Technologies, NAACL-HLT","volume":"1","author":"Devlin Jacob","year":"2019","unstructured":"Jacob Devlin , Ming-Wei Chang , Kenton Lee , and Kristina Toutanova . 2019 . BERT: Pre-training of Deep Bidirectional Transformers for Language Understanding . In Proceedings of the 2019 Conference of the North American Chapter of the Association for Computational Linguistics: Human Language Technologies, NAACL-HLT 2019, Volume 1 (Long and Short Papers). 4171--4186. Jacob Devlin, Ming-Wei Chang, Kenton Lee, and Kristina Toutanova. 2019. BERT: Pre-training of Deep Bidirectional Transformers for Language Understanding. In Proceedings of the 2019 Conference of the North American Chapter of the Association for Computational Linguistics: Human Language Technologies, NAACL-HLT 2019, Volume 1 (Long and Short Papers). 4171--4186."},{"key":"e_1_3_2_1_10_1","volume-title":"International Conference on Machine Learning. PMLR, 1126--1135","author":"Finn Chelsea","year":"2017","unstructured":"Chelsea Finn , Pieter Abbeel , and Sergey Levine . 2017 . Model-agnostic meta-learning for fast adaptation of deep networks . In International Conference on Machine Learning. PMLR, 1126--1135 . Chelsea Finn, Pieter Abbeel, and Sergey Levine. 2017. Model-agnostic meta-learning for fast adaptation of deep networks. In International Conference on Machine Learning. PMLR, 1126--1135."},{"key":"e_1_3_2_1_11_1","volume-title":"Proceedings of the 34th International Conference on Machine Learning, ICML 2017","volume":"70","author":"Gal Yarin","year":"2017","unstructured":"Yarin Gal , Riashat Islam , and Zoubin Ghahramani . 2017 . Deep Bayesian Active Learning with Image Data . In Proceedings of the 34th International Conference on Machine Learning, ICML 2017 , Vol. 70 . PMLR, 1183--1192. Yarin Gal, Riashat Islam, and Zoubin Ghahramani. 2017. Deep Bayesian Active Learning with Image Data. In Proceedings of the 34th International Conference on Machine Learning, ICML 2017, Vol. 70. PMLR, 1183--1192."},{"key":"e_1_3_2_1_12_1","volume-title":"Don't Stop Pretraining: Adapt Language Models to Domains and Tasks. arXiv preprint arXiv:2004.10964","author":"Gururangan Suchin","year":"2020","unstructured":"Suchin Gururangan , Ana Marasovi\u0107 , Swabha Swayamdipta , Kyle Lo , Iz Beltagy , Doug Downey , and Noah A Smith . 2020. Don't Stop Pretraining: Adapt Language Models to Domains and Tasks. arXiv preprint arXiv:2004.10964 ( 2020 ). Suchin Gururangan, Ana Marasovi\u0107, Swabha Swayamdipta, Kyle Lo, Iz Beltagy, Doug Downey, and Noah A Smith. 2020. Don't Stop Pretraining: Adapt Language Models to Domains and Tasks. arXiv preprint arXiv:2004.10964 (2020)."},{"key":"e_1_3_2_1_13_1","volume-title":"Revisiting Self-Training for Neural Sequence Generation. arxiv","author":"He Junxian","year":"1909","unstructured":"Junxian He , Jiatao Gu , Jiajun Shen , and Marc'Aurelio Ranzato . 2019. Revisiting Self-Training for Neural Sequence Generation. arxiv : 1909 .13788 [cs.LG] Junxian He, Jiatao Gu, Jiajun Shen, and Marc'Aurelio Ranzato. 2019. Revisiting Self-Training for Neural Sequence Generation. arxiv: 1909.13788 [cs.LG]"},{"key":"e_1_3_2_1_14_1","doi-asserted-by":"publisher","DOI":"10.1109\/TIT.1965.1053799"},{"key":"e_1_3_2_1_15_1","volume-title":"International Conference on Machine Learning. 2304--2313","author":"Jiang Lu","year":"2018","unstructured":"Lu Jiang , Zhengyuan Zhou , Thomas Leung , Li-Jia Li , and Li Fei-Fei . 2018 . Mentornet: Learning data-driven curriculum for very deep neural networks on corrupted labels . In International Conference on Machine Learning. 2304--2313 . Lu Jiang, Zhengyuan Zhou, Thomas Leung, Li-Jia Li, and Li Fei-Fei. 2018. Mentornet: Learning data-driven curriculum for very deep neural networks on corrupted labels. In International Conference on Machine Learning. 2304--2313."},{"key":"e_1_3_2_1_16_1","volume-title":"Self-Training with Weak Supervision. arXiv preprint arXiv:2104.05514","author":"Karamanolakis Giannis","year":"2021","unstructured":"Giannis Karamanolakis , Subhabrata Mukherjee , Guoqing Zheng , and Ahmed Hassan Awadallah . 2021. Self-Training with Weak Supervision. arXiv preprint arXiv:2104.05514 ( 2021 ). Giannis Karamanolakis, Subhabrata Mukherjee, Guoqing Zheng, and Ahmed Hassan Awadallah. 2021. Self-Training with Weak Supervision. arXiv preprint arXiv:2104.05514 (2021)."},{"key":"e_1_3_2_1_17_1","first-page":"2014","article-title":"Semi-supervised Learning with Deep Generative Models","volume":"27","author":"Kingma Diederik P.","year":"2014","unstructured":"Diederik P. Kingma , Shakir Mohamed , Danilo Jimenez Rezende , and Max Welling . 2014 . Semi-supervised Learning with Deep Generative Models . In Advances in Neural Information Processing Systems 27 , 2014 . 3581--3589. Diederik P. Kingma, Shakir Mohamed, Danilo Jimenez Rezende, and Max Welling. 2014. Semi-supervised Learning with Deep Generative Models. In Advances in Neural Information Processing Systems 27, 2014. 3581--3589.","journal-title":"Advances in Neural Information Processing Systems"},{"key":"e_1_3_2_1_18_1","volume-title":"Understanding black-box predictions via influence functions. arXiv preprint arXiv:1703.04730","author":"Koh Pang Wei","year":"2017","unstructured":"Pang Wei Koh and Percy Liang . 2017. Understanding black-box predictions via influence functions. arXiv preprint arXiv:1703.04730 ( 2017 ). Pang Wei Koh and Percy Liang. 2017. Understanding black-box predictions via influence functions. arXiv preprint arXiv:1703.04730 (2017)."},{"key":"e_1_3_2_1_19_1","unstructured":"Ksenia Konyushkova Raphael Sznitman and Pascal Fua. 2017. Learning active learning from data. In Advances in Neural Information Processing Systems.  Ksenia Konyushkova Raphael Sznitman and Pascal Fua. 2017. Learning active learning from data. In Advances in Neural Information Processing Systems."},{"key":"e_1_3_2_1_20_1","volume-title":"Understanding Self-Training for Gradual Domain Adaptation. arXiv preprint arXiv:2002.11361","author":"Kumar Ananya","year":"2020","unstructured":"Ananya Kumar , Tengyu Ma , and Percy Liang . 2020. Understanding Self-Training for Gradual Domain Adaptation. arXiv preprint arXiv:2002.11361 ( 2020 ). Ananya Kumar, Tengyu Ma, and Percy Liang. 2020. Understanding Self-Training for Gradual Domain Adaptation. arXiv preprint arXiv:2002.11361 (2020)."},{"key":"e_1_3_2_1_21_1","volume-title":"Advances in Neural Information Processing Systems 23. Curran Associates","author":"Kumar M. P.","unstructured":"M. P. Kumar , Benjamin Packer , and Daphne Koller . 2010. Self-Paced Learning for Latent Variable Models . In Advances in Neural Information Processing Systems 23. Curran Associates , Inc ., 1189--1197. M. P. Kumar, Benjamin Packer, and Daphne Koller. 2010. Self-Paced Learning for Latent Variable Models. In Advances in Neural Information Processing Systems 23. Curran Associates, Inc., 1189--1197."},{"key":"e_1_3_2_1_22_1","volume-title":"Building machines that learn and think like people. Behavioral and brain sciences","author":"Lake Brenden M","year":"2017","unstructured":"Brenden M Lake , Tomer D Ullman , Joshua B Tenenbaum , and Samuel J Gershman . 2017. Building machines that learn and think like people. Behavioral and brain sciences , Vol. 40 ( 2017 ). Brenden M Lake, Tomer D Ullman, Joshua B Tenenbaum, and Samuel J Gershman. 2017. Building machines that learn and think like people. Behavioral and brain sciences, Vol. 40 (2017)."},{"key":"e_1_3_2_1_23_1","unstructured":"Xinzhe Li Qianru Sun Yaoyao Liu Qin Zhou Shibao Zheng Tat-Seng Chua and Bernt Schiele. 2019. Learning to Self-Train for Semi-Supervised Few-Shot Classification. In Advances in Neural Information Processing Systems 32.  Xinzhe Li Qianru Sun Yaoyao Liu Qin Zhou Shibao Zheng Tat-Seng Chua and Bernt Schiele. 2019. Learning to Self-Train for Semi-Supervised Few-Shot Classification. In Advances in Neural Information Processing Systems 32."},{"key":"e_1_3_2_1_24_1","doi-asserted-by":"publisher","DOI":"10.1145\/3394486.3403149"},{"key":"e_1_3_2_1_25_1","doi-asserted-by":"publisher","DOI":"10.1109\/ICASSP.2013.6639301"},{"key":"e_1_3_2_1_26_1","volume-title":"RoBERTa: A Robustly Optimized BERT Pretraining Approach. CoRR","author":"Liu Yinhan","year":"2019","unstructured":"Yinhan Liu , Myle Ott , Naman Goyal , Jingfei Du , Mandar Joshi , Danqi Chen , Omer Levy , Mike Lewis , Luke Zettlemoyer , and Veselin Stoyanov . 2019. RoBERTa: A Robustly Optimized BERT Pretraining Approach. CoRR , Vol. abs\/ 1907 .11692 ( 2019 ). arxiv: 1907.11692 Yinhan Liu, Myle Ott, Naman Goyal, Jingfei Du, Mandar Joshi, Danqi Chen, Omer Levy, Mike Lewis, Luke Zettlemoyer, and Veselin Stoyanov. 2019. RoBERTa: A Robustly Optimized BERT Pretraining Approach. CoRR, Vol. abs\/1907.11692 (2019). arxiv: 1907.11692"},{"key":"e_1_3_2_1_27_1","volume-title":"Self-Paced Co-training. In Proceedings of the 34th International Conference on Machine Learning (Proceedings of Machine Learning Research","volume":"2284","author":"Ma Fan","year":"2017","unstructured":"Fan Ma , Deyu Meng , Qi Xie , Zina Li , and Xuanyi Dong . 2017 . Self-Paced Co-training. In Proceedings of the 34th International Conference on Machine Learning (Proceedings of Machine Learning Research , Vol. 70). 2275-- 2284 . Fan Ma, Deyu Meng, Qi Xie, Zina Li, and Xuanyi Dong. 2017. Self-Paced Co-training. In Proceedings of the 34th International Conference on Machine Learning (Proceedings of Machine Learning Research, Vol. 70). 2275--2284."},{"key":"e_1_3_2_1_28_1","volume-title":"Proceedings of the Human Language Technology Conference of the North American Chapter of the Association for Computational Linguistics: HLT-NAACL","author":"Miller Scott","year":"2004","unstructured":"Scott Miller , Jethran Guinness , and Alex Zamanian . 2004 . Name tagging with word clusters and discriminative training . In Proceedings of the Human Language Technology Conference of the North American Chapter of the Association for Computational Linguistics: HLT-NAACL 2004. 337--342. Scott Miller, Jethran Guinness, and Alex Zamanian. 2004. Name tagging with word clusters and discriminative training. In Proceedings of the Human Language Technology Conference of the North American Chapter of the Association for Computational Linguistics: HLT-NAACL 2004. 337--342."},{"key":"e_1_3_2_1_29_1","volume-title":"Virtual adversarial training: a regularization method for supervised and semi-supervised learning","author":"Miyato Takeru","year":"2018","unstructured":"Takeru Miyato , Shin-ichi Maeda, Masanori Koyama , and Shin Ishii . 2018. Virtual adversarial training: a regularization method for supervised and semi-supervised learning . IEEE transactions on pattern analysis and machine intelligence, Vol. 41 ( 2018 ). Takeru Miyato, Shin-ichi Maeda, Masanori Koyama, and Shin Ishii. 2018. Virtual adversarial training: a regularization method for supervised and semi-supervised learning. IEEE transactions on pattern analysis and machine intelligence, Vol. 41 (2018)."},{"key":"e_1_3_2_1_30_1","volume-title":"Uncertainty-aware self-training for few-shot text classification. Advances in Neural Information Processing Systems","author":"Mukherjee Subhabrata","year":"2020","unstructured":"Subhabrata Mukherjee and Ahmed Awadallah . 2020. Uncertainty-aware self-training for few-shot text classification. Advances in Neural Information Processing Systems ( 2020 ). Subhabrata Mukherjee and Ahmed Awadallah. 2020. Uncertainty-aware self-training for few-shot text classification. Advances in Neural Information Processing Systems (2020)."},{"key":"e_1_3_2_1_31_1","doi-asserted-by":"publisher","DOI":"10.18653\/v1\/P17-1178"},{"key":"e_1_3_2_1_32_1","unstructured":"Emmeleia Panagiota Mastoropoulou. 2019. Enhancing Deep Active Learning Using Selective Self-Training For Image Classification. Master's thesis. KTH School of Electrical Engineering and Computer Science (EECS).  Emmeleia Panagiota Mastoropoulou. 2019. Enhancing Deep Active Learning Using Selective Self-Training For Image Classification. Master's thesis. KTH School of Electrical Engineering and Computer Science (EECS)."},{"key":"e_1_3_2_1_33_1","volume-title":"Semi-supervised sequence tagging with bidirectional language models. arXiv preprint arXiv:1705.00108","author":"Peters Matthew E","year":"2017","unstructured":"Matthew E Peters , Waleed Ammar , Chandra Bhagavatula , and Russell Power . 2017. Semi-supervised sequence tagging with bidirectional language models. arXiv preprint arXiv:1705.00108 ( 2017 ). Matthew E Peters, Waleed Ammar, Chandra Bhagavatula, and Russell Power. 2017. Semi-supervised sequence tagging with bidirectional language models. arXiv preprint arXiv:1705.00108 (2017)."},{"key":"e_1_3_2_1_34_1","unstructured":"Slav Petrov and Ryan McDonald. 2012. Overview of the 2012 shared task on parsing the web. (2012).  Slav Petrov and Ryan McDonald. 2012. Overview of the 2012 shared task on parsing the web. (2012)."},{"key":"e_1_3_2_1_35_1","unstructured":"Alec Radford Jeffrey Wu Rewon Child David Luan Dario Amodei and Ilya Sutskever. 2019. Language Models are Unsupervised Multitask Learners. (2019).  Alec Radford Jeffrey Wu Rewon Child David Luan Dario Amodei and Ilya Sutskever. 2019. Language Models are Unsupervised Multitask Learners. (2019)."},{"key":"e_1_3_2_1_36_1","volume-title":"International Conference on Machine Learning. 4334--4343","author":"Ren Mengye","year":"2018","unstructured":"Mengye Ren , Wenyuan Zeng , Bin Yang , and Raquel Urtasun . 2018 . Learning to Reweight Examples for Robust Deep Learning . In International Conference on Machine Learning. 4334--4343 . Mengye Ren, Wenyuan Zeng, Bin Yang, and Raquel Urtasun. 2018. Learning to Reweight Examples for Robust Deep Learning. In International Conference on Machine Learning. 4334--4343."},{"key":"e_1_3_2_1_37_1","doi-asserted-by":"publisher","DOI":"10.18653\/v1\/P18-1096"},{"key":"e_1_3_2_1_38_1","volume-title":"Seventh Conference on Natural Language Learning at HLT-NAACL","author":"Erik","year":"2003","unstructured":"Erik F. Tjong Kim Sang and Fien De Meulder. [n.d.]. Introduction to the CoNLL-2003 Shared Task: Language-Independent Named Entity Recognition . In Seventh Conference on Natural Language Learning at HLT-NAACL 2003 . Erik F. Tjong Kim Sang and Fien De Meulder. [n.d.]. Introduction to the CoNLL-2003 Shared Task: Language-Independent Named Entity Recognition. In Seventh Conference on Natural Language Learning at HLT-NAACL 2003."},{"key":"e_1_3_2_1_39_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR.2016.89"},{"key":"e_1_3_2_1_40_1","volume-title":"5th International Conference on Learning Representations, ICLR","author":"Tarvainen Antti","year":"2017","unstructured":"Antti Tarvainen and Harri Valpola . 2017 . Mean teachers are better role models: Weight-averaged consistency targets improve semi-supervised deep learning results . In 5th International Conference on Learning Representations, ICLR 2017. Antti Tarvainen and Harri Valpola. 2017. Mean teachers are better role models: Weight-averaged consistency targets improve semi-supervised deep learning results. In 5th International Conference on Learning Representations, ICLR 2017."},{"key":"e_1_3_2_1_41_1","doi-asserted-by":"crossref","unstructured":"Sebastian Thrun and Lorien Pratt. 1998. Learning to learn: Introduction and overview. In Learning to learn. 3--17.  Sebastian Thrun and Lorien Pratt. 1998. Learning to learn: Introduction and overview. In Learning to learn. 3--17.","DOI":"10.1007\/978-1-4615-5529-2_1"},{"key":"e_1_3_2_1_42_1","volume-title":"Representing Text Chunks. In Ninth Conference of the European Chapter of the Association for Computational Linguistics.","author":"Tjong Erik F","year":"1999","unstructured":"Erik F Tjong , Kim Sang , and Jorn Veenstra . 1999 . Representing Text Chunks. In Ninth Conference of the European Chapter of the Association for Computational Linguistics. Erik F Tjong, Kim Sang, and Jorn Veenstra. 1999. Representing Text Chunks. In Ninth Conference of the European Chapter of the Association for Computational Linguistics."},{"key":"e_1_3_2_1_43_1","doi-asserted-by":"publisher","DOI":"10.1145\/3394486.3403303"},{"key":"e_1_3_2_1_44_1","doi-asserted-by":"publisher","DOI":"10.1609\/aaai.v34i01.5389"},{"key":"e_1_3_2_1_45_1","volume-title":"Proceedings of the IEEE\/CVF Conference on Computer Vision and Pattern Recognition (CVPR).","author":"Xie Qizhe","unstructured":"Qizhe Xie , Minh-Thang Luong , Eduard Hovy , and Quoc V. Le . 2020. Self-Training With Noisy Student Improves ImageNet Classification . In Proceedings of the IEEE\/CVF Conference on Computer Vision and Pattern Recognition (CVPR). Qizhe Xie, Minh-Thang Luong, Eduard Hovy, and Quoc V. Le. 2020. Self-Training With Noisy Student Improves ImageNet Classification. In Proceedings of the IEEE\/CVF Conference on Computer Vision and Pattern Recognition (CVPR)."},{"key":"e_1_3_2_1_46_1","volume-title":"5th International Conference on Learning Representations, ICLR","author":"Zhang Chiyuan","year":"2017","unstructured":"Chiyuan Zhang , Samy Bengio , Moritz Hardt , Benjamin Recht , and Oriol Vinyals . 2017 . Understanding deep learning requires rethinking generalization . In 5th International Conference on Learning Representations, ICLR 2017. Chiyuan Zhang, Samy Bengio, Moritz Hardt, Benjamin Recht, and Oriol Vinyals. 2017. Understanding deep learning requires rethinking generalization. In 5th International Conference on Learning Representations, ICLR 2017."},{"key":"e_1_3_2_1_47_1","volume-title":"Ekin Dogus Cubuk, and Quoc Le","author":"Zoph Barret","year":"2020","unstructured":"Barret Zoph , Golnaz Ghiasi , Tsung-Yi Lin , Yin Cui , Hanxiao Liu , Ekin Dogus Cubuk, and Quoc Le . 2020 . Rethinking pre-training and self-training. Advances in Neural Information Processing Systems , Vol. 33 (2020). Barret Zoph, Golnaz Ghiasi, Tsung-Yi Lin, Yin Cui, Hanxiao Liu, Ekin Dogus Cubuk, and Quoc Le. 2020. Rethinking pre-training and self-training. Advances in Neural Information Processing Systems, Vol. 33 (2020)."}],"event":{"name":"KDD '21: The 27th ACM SIGKDD Conference on Knowledge Discovery and Data Mining","location":"Virtual Event Singapore","acronym":"KDD '21","sponsor":["SIGMOD ACM Special Interest Group on Management of Data","SIGKDD ACM Special Interest Group on Knowledge Discovery in Data"]},"container-title":["Proceedings of the 27th ACM SIGKDD Conference on Knowledge Discovery &amp; Data Mining"],"original-title":[],"link":[{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/3447548.3467235","content-type":"unspecified","content-version":"vor","intended-application":"text-mining"},{"URL":"https:\/\/dl.acm.org\/doi\/abs\/10.1145\/3447548.3467235","content-type":"text\/html","content-version":"vor","intended-application":"syndication"},{"URL":"https:\/\/dl.acm.org\/doi\/pdf\/10.1145\/3447548.3467235","content-type":"application\/pdf","content-version":"vor","intended-application":"syndication"},{"URL":"https:\/\/dl.acm.org\/doi\/pdf\/10.1145\/3447548.3467235","content-type":"unspecified","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2025,6,17]],"date-time":"2025-06-17T20:18:28Z","timestamp":1750191508000},"score":1,"resource":{"primary":{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/3447548.3467235"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2021,8,14]]},"references-count":47,"alternative-id":["10.1145\/3447548.3467235","10.1145\/3447548"],"URL":"https:\/\/doi.org\/10.1145\/3447548.3467235","relation":{},"subject":[],"published":{"date-parts":[[2021,8,14]]},"assertion":[{"value":"2021-08-14","order":2,"name":"published","label":"Published","group":{"name":"publication_history","label":"Publication History"}}]}}