{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2026,4,30]],"date-time":"2026-04-30T16:57:39Z","timestamp":1777568259953,"version":"3.51.4"},"publisher-location":"New York, NY, USA","reference-count":47,"publisher":"ACM","license":[{"start":{"date-parts":[[2022,10,10]],"date-time":"2022-10-10T00:00:00Z","timestamp":1665360000000},"content-version":"vor","delay-in-days":0,"URL":"https:\/\/www.acm.org\/publications\/policies\/copyright_policy#Background"}],"funder":[{"name":"the Science and Technology Commission of Shanghai Municipality","award":["20DZ2220400"],"award-info":[{"award-number":["20DZ2220400"]}]},{"name":"NSFC","award":["62176156"],"award-info":[{"award-number":["62176156"]}]}],"content-domain":{"domain":["dl.acm.org"],"crossmark-restriction":true},"short-container-title":[],"published-print":{"date-parts":[[2022,10,10]]},"DOI":"10.1145\/3503161.3547980","type":"proceedings-article","created":{"date-parts":[[2022,10,10]],"date-time":"2022-10-10T15:43:01Z","timestamp":1665416581000},"page":"2124-2134","update-policy":"https:\/\/doi.org\/10.1145\/crossmark-policy","source":"Crossref","is-referenced-by-count":9,"title":["Semi-supervised Learning for Multi-label Video Action Detection"],"prefix":"10.1145","author":[{"given":"Hongcheng","family":"Zhang","sequence":"first","affiliation":[{"name":"Shanghai Jiao Tong University, Shanghai, China"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Xu","family":"Zhao","sequence":"additional","affiliation":[{"name":"Shanghai Jiao Tong University, Shanghai, China"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Dongqi","family":"Wang","sequence":"additional","affiliation":[{"name":"Shanghai Jiao Tong University, Shanghai, China"}],"role":[{"role":"author","vocabulary":"crossref"}]}],"member":"320","published-online":{"date-parts":[[2022,10,10]]},"reference":[{"key":"e_1_3_2_2_1_1","volume-title":"Learning with pseudo-ensembles. Advances in neural information processing systems","author":"Bachman Philip","year":"2014","unstructured":"Philip Bachman , Ouais Alsharif , and Doina Precup . 2014. Learning with pseudo-ensembles. Advances in neural information processing systems , Vol. 27 ( 2014 ). Philip Bachman, Ouais Alsharif, and Doina Precup. 2014. Learning with pseudo-ensembles. Advances in neural information processing systems , Vol. 27 (2014)."},{"key":"e_1_3_2_2_2_1","volume-title":"Remixmatch: Semi-supervised learning with distribution alignment and augmentation anchoring. arXiv preprint arXiv:1911.09785","author":"Berthelot David","year":"2019","unstructured":"David Berthelot , Nicholas Carlini , Ekin D Cubuk , Alex Kurakin , Kihyuk Sohn , Han Zhang , and Colin Raffel . 2019 a. Remixmatch: Semi-supervised learning with distribution alignment and augmentation anchoring. arXiv preprint arXiv:1911.09785 (2019). David Berthelot, Nicholas Carlini, Ekin D Cubuk, Alex Kurakin, Kihyuk Sohn, Han Zhang, and Colin Raffel. 2019a. Remixmatch: Semi-supervised learning with distribution alignment and augmentation anchoring. arXiv preprint arXiv:1911.09785 (2019)."},{"key":"e_1_3_2_2_3_1","volume-title":"Advances in Neural Information Processing Systems","volume":"32","author":"Berthelot David","year":"2019","unstructured":"David Berthelot , Nicholas Carlini , Ian Goodfellow , Nicolas Papernot , Avital Oliver , and Colin A Raffel . 2019 b. Mixmatch: A holistic approach to semi-supervised learning . Advances in Neural Information Processing Systems , Vol. 32 (2019). David Berthelot, Nicholas Carlini, Ian Goodfellow, Nicolas Papernot, Avital Oliver, and Colin A Raffel. 2019b. Mixmatch: A holistic approach to semi-supervised learning. Advances in Neural Information Processing Systems , Vol. 32 (2019)."},{"key":"e_1_3_2_2_4_1","doi-asserted-by":"publisher","DOI":"10.1016\/j.neunet.2018.07.011"},{"key":"e_1_3_2_2_5_1","volume-title":"International Conference on Machine Learning. PMLR, 872--881","author":"Byrd Jonathon","year":"2019","unstructured":"Jonathon Byrd and Zachary Lipton . 2019 . What is the effect of importance weighting in deep learning? . In International Conference on Machine Learning. PMLR, 872--881 . Jonathon Byrd and Zachary Lipton. 2019. What is the effect of importance weighting in deep learning?. In International Conference on Machine Learning. PMLR, 872--881."},{"key":"e_1_3_2_2_6_1","volume-title":"Learning imbalanced datasets with label-distribution-aware margin loss. Advances in neural information processing systems","author":"Cao Kaidi","year":"2019","unstructured":"Kaidi Cao , Colin Wei , Adrien Gaidon , Nikos Arechiga , and Tengyu Ma. 2019. Learning imbalanced datasets with label-distribution-aware margin loss. Advances in neural information processing systems , Vol. 32 ( 2019 ). Kaidi Cao, Colin Wei, Adrien Gaidon, Nikos Arechiga, and Tengyu Ma. 2019. Learning imbalanced datasets with label-distribution-aware margin loss. Advances in neural information processing systems , Vol. 32 (2019)."},{"key":"e_1_3_2_2_7_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR.2017.502"},{"key":"e_1_3_2_2_8_1","doi-asserted-by":"publisher","DOI":"10.1109\/ICCV48922.2021.00807"},{"key":"e_1_3_2_2_9_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPRW50498.2020.00359"},{"key":"e_1_3_2_2_10_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR.2019.00949"},{"key":"e_1_3_2_2_11_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR.2009.5206848"},{"key":"e_1_3_2_2_12_1","doi-asserted-by":"publisher","DOI":"10.1109\/ICCV.2019.00630"},{"key":"e_1_3_2_2_13_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR.2019.00033"},{"key":"e_1_3_2_2_14_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR.2018.00633"},{"key":"e_1_3_2_2_15_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR46437.2021.01484"},{"key":"e_1_3_2_2_16_1","doi-asserted-by":"publisher","DOI":"10.1109\/ICCV.2017.620"},{"key":"e_1_3_2_2_17_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR.2016.580"},{"key":"e_1_3_2_2_18_1","doi-asserted-by":"publisher","DOI":"10.1109\/ICCV.2017.472"},{"key":"e_1_3_2_2_19_1","volume-title":"Decoupling representation and classifier for long-tailed recognition. arXiv preprint arXiv:1910.09217","author":"Kang Bingyi","year":"2019","unstructured":"Bingyi Kang , Saining Xie , Marcus Rohrbach , Zhicheng Yan , Albert Gordo , Jiashi Feng , and Yannis Kalantidis . 2019. Decoupling representation and classifier for long-tailed recognition. arXiv preprint arXiv:1910.09217 ( 2019 ). Bingyi Kang, Saining Xie, Marcus Rohrbach, Zhicheng Yan, Albert Gordo, Jiashi Feng, and Yannis Kalantidis. 2019. Decoupling representation and classifier for long-tailed recognition. arXiv preprint arXiv:1910.09217 (2019)."},{"key":"e_1_3_2_2_20_1","volume-title":"You only watch once: A unified cnn architecture for real-time spatiotemporal action localization. arXiv preprint arXiv:1911.06644","author":"Okan K\u00f6p\u00fckl\u00fc","year":"2019","unstructured":"Okan K\u00f6p\u00fckl\u00fc , Xiangyu Wei , and Gerhard Rigoll . 2019. You only watch once: A unified cnn architecture for real-time spatiotemporal action localization. arXiv preprint arXiv:1911.06644 ( 2019 ). Okan K\u00f6p\u00fckl\u00fc , Xiangyu Wei, and Gerhard Rigoll. 2019. You only watch once: A unified cnn architecture for real-time spatiotemporal action localization. arXiv preprint arXiv:1911.06644 (2019)."},{"key":"e_1_3_2_2_21_1","volume-title":"End-to-End Semi-Supervised Learning for Video Action Detection. arXiv preprint arXiv:2203.04251","author":"Kumar Akash","year":"2022","unstructured":"Akash Kumar and Yogesh Singh Rawat . 2022. End-to-End Semi-Supervised Learning for Video Action Detection. arXiv preprint arXiv:2203.04251 ( 2022 ). Akash Kumar and Yogesh Singh Rawat. 2022. End-to-End Semi-Supervised Learning for Video Action Detection. arXiv preprint arXiv:2203.04251 (2022)."},{"key":"e_1_3_2_2_22_1","volume-title":"Temporal ensembling for semi-supervised learning. arXiv preprint arXiv:1610.02242","author":"Laine Samuli","year":"2016","unstructured":"Samuli Laine and Timo Aila . 2016. Temporal ensembling for semi-supervised learning. arXiv preprint arXiv:1610.02242 ( 2016 ). Samuli Laine and Timo Aila. 2016. Temporal ensembling for semi-supervised learning. arXiv preprint arXiv:1610.02242 (2016)."},{"key":"e_1_3_2_2_23_1","volume-title":"Workshop on challenges in representation learning, ICML","volume":"3","author":"Dong-Hyun","unstructured":"Dong-Hyun Lee et al. 2013. Pseudo-label: The simple and efficient semi-supervised learning method for deep neural networks . In Workshop on challenges in representation learning, ICML , Vol. 3 . 896. Dong-Hyun Lee et al. 2013. Pseudo-label: The simple and efficient semi-supervised learning method for deep neural networks. In Workshop on challenges in representation learning, ICML, Vol. 3. 896."},{"key":"e_1_3_2_2_24_1","volume-title":"Alexander Vostrikov, and Andrew Zisserman.","author":"Li Ang","year":"2020","unstructured":"Ang Li , Meghana Thotakuri , David A Ross , Jo ao Carreira , Alexander Vostrikov, and Andrew Zisserman. 2020 a. The ava-kinetics localized human actions video dataset. arXiv preprint arXiv:2005.00214 (2020). Ang Li, Meghana Thotakuri, David A Ross, Jo ao Carreira, Alexander Vostrikov, and Andrew Zisserman. 2020a. The ava-kinetics localized human actions video dataset. arXiv preprint arXiv:2005.00214 (2020)."},{"key":"e_1_3_2_2_25_1","doi-asserted-by":"publisher","DOI":"10.1007\/978-3-030-58517-4_5"},{"key":"e_1_3_2_2_26_1","doi-asserted-by":"publisher","DOI":"10.1145\/3474085.3475374"},{"key":"e_1_3_2_2_27_1","doi-asserted-by":"publisher","DOI":"10.1007\/978-3-319-10602-1_48"},{"key":"e_1_3_2_2_28_1","doi-asserted-by":"publisher","DOI":"10.1145\/3474085.3475503"},{"key":"e_1_3_2_2_29_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR46437.2021.00053"},{"key":"e_1_3_2_2_30_1","volume-title":"Faster r-cnn: Towards real-time object detection with region proposal networks. arXiv preprint arXiv:1506.01497","author":"Ren Shaoqing","year":"2015","unstructured":"Shaoqing Ren , Kaiming He , Ross Girshick , and Jian Sun . 2015. Faster r-cnn: Towards real-time object detection with region proposal networks. arXiv preprint arXiv:1506.01497 ( 2015 ). Shaoqing Ren, Kaiming He, Ross Girshick, and Jian Sun. 2015. Faster r-cnn: Towards real-time object detection with region proposal networks. arXiv preprint arXiv:1506.01497 (2015)."},{"key":"e_1_3_2_2_31_1","doi-asserted-by":"publisher","DOI":"10.1007\/978-3-319-46478-7_29"},{"key":"e_1_3_2_2_32_1","volume-title":"Super-convergence: Very fast training of neural networks using large learning rates. In Artificial intelligence and machine learning for multi-domain operations applications","author":"Smith Leslie N","year":"2019","unstructured":"Leslie N Smith and Nicholay Topin . 2019 . Super-convergence: Very fast training of neural networks using large learning rates. In Artificial intelligence and machine learning for multi-domain operations applications , Vol. 11006 . International Society for Optics and Photonics , 1100612. Leslie N Smith and Nicholay Topin. 2019. Super-convergence: Very fast training of neural networks using large learning rates. In Artificial intelligence and machine learning for multi-domain operations applications, Vol. 11006. International Society for Optics and Photonics, 1100612."},{"key":"e_1_3_2_2_33_1","first-page":"596","article-title":"Fixmatch: Simplifying semi-supervised learning with consistency and confidence","volume":"33","author":"Sohn Kihyuk","year":"2020","unstructured":"Kihyuk Sohn , David Berthelot , Nicholas Carlini , Zizhao Zhang , Han Zhang , Colin A Raffel , Ekin Dogus Cubuk , Alexey Kurakin , and Chun-Liang Li . 2020 . Fixmatch: Simplifying semi-supervised learning with consistency and confidence . Advances in Neural Information Processing Systems , Vol. 33 (2020), 596 -- 608 . Kihyuk Sohn, David Berthelot, Nicholas Carlini, Zizhao Zhang, Han Zhang, Colin A Raffel, Ekin Dogus Cubuk, Alexey Kurakin, and Chun-Liang Li. 2020. Fixmatch: Simplifying semi-supervised learning with consistency and confidence. Advances in Neural Information Processing Systems , Vol. 33 (2020), 596--608.","journal-title":"Advances in Neural Information Processing Systems"},{"key":"e_1_3_2_2_34_1","volume-title":"Amir Roshan Zamir, and Mubarak Shah","author":"Soomro Khurram","year":"2012","unstructured":"Khurram Soomro , Amir Roshan Zamir, and Mubarak Shah . 2012 . UCF101: A dataset of 101 human actions classes from videos in the wild. arXiv preprint arXiv:1212.0402 (2012). Khurram Soomro, Amir Roshan Zamir, and Mubarak Shah. 2012. UCF101: A dataset of 101 human actions classes from videos in the wild. arXiv preprint arXiv:1212.0402 (2012)."},{"key":"e_1_3_2_2_35_1","doi-asserted-by":"publisher","DOI":"10.1007\/978-3-030-58555-6_5"},{"key":"e_1_3_2_2_36_1","volume-title":"Mean teachers are better role models: Weight-averaged consistency targets improve semi-supervised deep learning results. Advances in neural information processing systems","author":"Tarvainen Antti","year":"2017","unstructured":"Antti Tarvainen and Harri Valpola . 2017. Mean teachers are better role models: Weight-averaged consistency targets improve semi-supervised deep learning results. Advances in neural information processing systems , Vol. 30 ( 2017 ). Antti Tarvainen and Harri Valpola. 2017. Mean teachers are better role models: Weight-averaged consistency targets improve semi-supervised deep learning results. Advances in neural information processing systems , Vol. 30 (2017)."},{"key":"e_1_3_2_2_37_1","volume-title":"Advances in Neural Information Processing Systems","volume":"30","author":"Wang Yu-Xiong","year":"2017","unstructured":"Yu-Xiong Wang , Deva Ramanan , and Martial Hebert . 2017 . Learning to model the tail . Advances in Neural Information Processing Systems , Vol. 30 (2017). Yu-Xiong Wang, Deva Ramanan, and Martial Hebert. 2017. Learning to model the tail. Advances in Neural Information Processing Systems , Vol. 30 (2017)."},{"key":"e_1_3_2_2_38_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR.2019.00037"},{"key":"e_1_3_2_2_39_1","doi-asserted-by":"publisher","DOI":"10.1007\/978-3-030-58595-2_27"},{"key":"e_1_3_2_2_40_1","doi-asserted-by":"publisher","DOI":"10.1007\/978-3-030-58548-8_10"},{"key":"e_1_3_2_2_41_1","first-page":"6256","article-title":"Unsupervised data augmentation for consistency training","volume":"33","author":"Xie Qizhe","year":"2020","unstructured":"Qizhe Xie , Zihang Dai , Eduard Hovy , Thang Luong , and Quoc Le . 2020 a. Unsupervised data augmentation for consistency training . Advances in Neural Information Processing Systems , Vol. 33 (2020), 6256 -- 6268 . Qizhe Xie, Zihang Dai, Eduard Hovy, Thang Luong, and Quoc Le. 2020a. Unsupervised data augmentation for consistency training. Advances in Neural Information Processing Systems , Vol. 33 (2020), 6256--6268.","journal-title":"Advances in Neural Information Processing Systems"},{"key":"e_1_3_2_2_42_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR42600.2020.01070"},{"key":"e_1_3_2_2_43_1","doi-asserted-by":"publisher","DOI":"10.1007\/978-3-030-01267-0_19"},{"key":"e_1_3_2_2_44_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR.2019.00035"},{"key":"e_1_3_2_2_45_1","first-page":"19290","article-title":"Rethinking the value of labels for improving class-imbalanced learning","volume":"33","author":"Yang Yuzhe","year":"2020","unstructured":"Yuzhe Yang and Zhi Xu . 2020 . Rethinking the value of labels for improving class-imbalanced learning . Advances in Neural Information Processing Systems , Vol. 33 (2020), 19290 -- 19301 . Yuzhe Yang and Zhi Xu. 2020. Rethinking the value of labels for improving class-imbalanced learning. Advances in Neural Information Processing Systems , Vol. 33 (2020), 19290--19301.","journal-title":"Advances in Neural Information Processing Systems"},{"key":"e_1_3_2_2_46_1","volume-title":"Advances in Neural Information Processing Systems","volume":"34","author":"Zhang Bowen","year":"2021","unstructured":"Bowen Zhang , Yidong Wang , Wenxin Hou , Hao Wu , Jindong Wang , Manabu Okumura , and Takahiro Shinozaki . 2021 . Flexmatch: Boosting semi-supervised learning with curriculum pseudo labeling . Advances in Neural Information Processing Systems , Vol. 34 (2021). Bowen Zhang, Yidong Wang, Wenxin Hou, Hao Wu, Jindong Wang, Manabu Okumura, and Takahiro Shinozaki. 2021. Flexmatch: Boosting semi-supervised learning with curriculum pseudo labeling. Advances in Neural Information Processing Systems , Vol. 34 (2021)."},{"key":"e_1_3_2_2_47_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR42600.2020.00974"}],"event":{"name":"MM '22: The 30th ACM International Conference on Multimedia","location":"Lisboa Portugal","acronym":"MM '22","sponsor":["SIGMM ACM Special Interest Group on Multimedia"]},"container-title":["Proceedings of the 30th ACM International Conference on Multimedia"],"original-title":[],"link":[{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/3503161.3547980","content-type":"unspecified","content-version":"vor","intended-application":"text-mining"},{"URL":"https:\/\/dl.acm.org\/doi\/pdf\/10.1145\/3503161.3547980","content-type":"unspecified","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2025,6,17]],"date-time":"2025-06-17T19:00:31Z","timestamp":1750186831000},"score":1,"resource":{"primary":{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/3503161.3547980"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2022,10,10]]},"references-count":47,"alternative-id":["10.1145\/3503161.3547980","10.1145\/3503161"],"URL":"https:\/\/doi.org\/10.1145\/3503161.3547980","relation":{},"subject":[],"published":{"date-parts":[[2022,10,10]]},"assertion":[{"value":"2022-10-10","order":2,"name":"published","label":"Published","group":{"name":"publication_history","label":"Publication History"}}]}}