{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2025,10,1]],"date-time":"2025-10-01T15:27:54Z","timestamp":1759332474201,"version":"3.41.0"},"publisher-location":"New York, NY, USA","reference-count":68,"publisher":"ACM","license":[{"start":{"date-parts":[[2022,10,10]],"date-time":"2022-10-10T00:00:00Z","timestamp":1665360000000},"content-version":"vor","delay-in-days":0,"URL":"https:\/\/www.acm.org\/publications\/policies\/copyright_policy#Background"}],"funder":[{"name":"National Key R&D Program of China","award":["No.2020YFC0832505"],"award-info":[{"award-number":["No.2020YFC0832505"]}]},{"name":"National Natural Science Foundation of China","award":["No.61836002"],"award-info":[{"award-number":["No.61836002"]}]},{"name":"Zhejiang Natural Science Foundation","award":["LR19F020006"],"award-info":[{"award-number":["LR19F020006"]}]}],"content-domain":{"domain":["dl.acm.org"],"crossmark-restriction":true},"short-container-title":[],"published-print":{"date-parts":[[2022,10,10]]},"DOI":"10.1145\/3503161.3547794","type":"proceedings-article","created":{"date-parts":[[2022,10,10]],"date-time":"2022-10-10T15:43:01Z","timestamp":1665416581000},"page":"6125-6135","update-policy":"https:\/\/doi.org\/10.1145\/crossmark-policy","source":"Crossref","is-referenced-by-count":1,"title":["Set-Based Face Recognition Beyond Disentanglement: Burstiness Suppression With Variance Vocabulary"],"prefix":"10.1145","author":[{"given":"Jiong","family":"Wang","sequence":"first","affiliation":[{"name":"Zhejiang University, Hangzhou, China"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Zhou","family":"Zhao","sequence":"additional","affiliation":[{"name":"Zhejiang University, Hangzhou, China"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Fei","family":"Wu","sequence":"additional","affiliation":[{"name":"Zhejiang University, Hangzhou, China"}],"role":[{"role":"author","vocabulary":"crossref"}]}],"member":"320","published-online":{"date-parts":[[2022,10,10]]},"reference":[{"key":"e_1_3_2_2_1_1","doi-asserted-by":"publisher","DOI":"10.1016\/j.imavis.2005.08.002"},{"key":"e_1_3_2_2_2_1","volume-title":"Proceedings of the IEEE international conference on computer vision. 1269--1277","author":"Babenko Artem","year":"2015","unstructured":"Artem Babenko and Victor Lempitsky . 2015 . Aggregating local deep features for image retrieval . In Proceedings of the IEEE international conference on computer vision. 1269--1277 . Artem Babenko and Victor Lempitsky. 2015. Aggregating local deep features for image retrieval. In Proceedings of the IEEE international conference on computer vision. 1269--1277."},{"key":"e_1_3_2_2_3_1","doi-asserted-by":"publisher","DOI":"10.1109\/FG.2018.00020"},{"key":"e_1_3_2_2_4_1","doi-asserted-by":"publisher","DOI":"10.1109\/ICCV.2019.01022"},{"key":"e_1_3_2_2_5_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR.2010.5539965"},{"key":"e_1_3_2_2_6_1","volume-title":"Isolating sources of disentanglement in variational autoencoders. Advances in neural information processing systems","author":"Chen Ricky TQ","year":"2018","unstructured":"Ricky TQ Chen , Xuechen Li , Roger B Grosse , and David K Duvenaud . 2018. Isolating sources of disentanglement in variational autoencoders. Advances in neural information processing systems , Vol. 31 ( 2018 ). Ricky TQ Chen, Xuechen Li, Roger B Grosse, and David K Duvenaud. 2018. Isolating sources of disentanglement in variational autoencoders. Advances in neural information processing systems, Vol. 31 (2018)."},{"key":"e_1_3_2_2_7_1","doi-asserted-by":"publisher","DOI":"10.1145\/3474085.3475351"},{"key":"e_1_3_2_2_8_1","volume-title":"Mean shift, mode seeking, and clustering","author":"Cheng Yizong","year":"1995","unstructured":"Yizong Cheng . 1995. Mean shift, mode seeking, and clustering . IEEE transactions on pattern analysis and machine intelligence, Vol. 17 , 8 ( 1995 ), 790--799. Yizong Cheng. 1995. Mean shift, mode seeking, and clustering. IEEE transactions on pattern analysis and machine intelligence, Vol. 17, 8 (1995), 790--799."},{"key":"e_1_3_2_2_9_1","doi-asserted-by":"publisher","DOI":"10.1017\/S1351324900000139"},{"key":"e_1_3_2_2_10_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR.2019.00482"},{"key":"e_1_3_2_2_11_1","volume-title":"Disentangling factors of variation via generative entangling. arXiv preprint arXiv:1210.5474","author":"Desjardins Guillaume","year":"2012","unstructured":"Guillaume Desjardins , Aaron Courville , and Yoshua Bengio . 2012. Disentangling factors of variation via generative entangling. arXiv preprint arXiv:1210.5474 ( 2012 ). Guillaume Desjardins, Aaron Courville, and Yoshua Bengio. 2012. Disentangling factors of variation via generative entangling. arXiv preprint arXiv:1210.5474 (2012)."},{"key":"e_1_3_2_2_12_1","volume-title":"Advances in Neural Information Processing Systems","volume":"32","author":"Eom Chanho","year":"2019","unstructured":"Chanho Eom and Bumsub Ham . 2019 . Learning disentangled representation for robust person re-identification . Advances in Neural Information Processing Systems , Vol. 32 (2019). Chanho Eom and Bumsub Ham. 2019. Learning disentangled representation for robust person re-identification. Advances in Neural Information Processing Systems, Vol. 32 (2019)."},{"key":"e_1_3_2_2_13_1","volume-title":"Recurrent embedding aggregation network for video face recognition. arXiv preprint arXiv:1904.12019","author":"Gong Sixue","year":"2019","unstructured":"Sixue Gong , Yichun Shi , and Anil K Jain . 2019a. Recurrent embedding aggregation network for video face recognition. arXiv preprint arXiv:1904.12019 ( 2019 ). Sixue Gong, Yichun Shi, and Anil K Jain. 2019a. Recurrent embedding aggregation network for video face recognition. arXiv preprint arXiv:1904.12019 (2019)."},{"key":"e_1_3_2_2_14_1","doi-asserted-by":"publisher","DOI":"10.1109\/ICB45273.2019.8987385"},{"key":"e_1_3_2_2_15_1","doi-asserted-by":"publisher","DOI":"10.1007\/978-3-319-46487-9_6"},{"key":"e_1_3_2_2_16_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR.2016.90"},{"key":"e_1_3_2_2_17_1","doi-asserted-by":"publisher","DOI":"10.1109\/ICDM.2007.17"},{"key":"e_1_3_2_2_18_1","unstructured":"Irina Higgins Loic Matthey Arka Pal Christopher Burgess Xavier Glorot Matthew Botvinick Shakir Mohamed and Alexander Lerchner. 2016. beta-vae: Learning basic visual concepts with a constrained variational framework. (2016).  Irina Higgins Loic Matthey Arka Pal Christopher Burgess Xavier Glorot Matthew Botvinick Shakir Mohamed and Alexander Lerchner. 2016. beta-vae: Learning basic visual concepts with a constrained variational framework. (2016)."},{"key":"e_1_3_2_2_19_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR.2018.00745"},{"key":"e_1_3_2_2_20_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR.2015.7298609"},{"key":"e_1_3_2_2_21_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR.2009.5206609"},{"key":"e_1_3_2_2_22_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR.2014.417"},{"key":"e_1_3_2_2_23_1","volume-title":"International Conference on Machine Learning. PMLR, 2649--2658","author":"Kim Hyunjik","year":"2018","unstructured":"Hyunjik Kim and Andriy Mnih . 2018 . Disentangling by factorising . In International Conference on Machine Learning. PMLR, 2649--2658 . Hyunjik Kim and Andriy Mnih. 2018. Disentangling by factorising. In International Conference on Machine Learning. PMLR, 2649--2658."},{"key":"e_1_3_2_2_24_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR52688.2022.01819"},{"key":"e_1_3_2_2_25_1","volume-title":"Auto-encoding variational bayes. arXiv preprint arXiv:1312.6114","author":"Kingma Diederik P","year":"2013","unstructured":"Diederik P Kingma and Max Welling . 2013. Auto-encoding variational bayes. arXiv preprint arXiv:1312.6114 ( 2013 ). Diederik P Kingma and Max Welling. 2013. Auto-encoding variational bayes. arXiv preprint arXiv:1312.6114 (2013)."},{"key":"e_1_3_2_2_26_1","volume-title":"2003 IEEE Computer Society Conference on Computer Vision and Pattern Recognition, 2003. Proceedings.","volume":"1","author":"Lee Kuang-Chih","year":"2003","unstructured":"Kuang-Chih Lee , Jeffrey Ho , Ming-Hsuan Yang , and David Kriegman . 2003 . Video-based face recognition using probabilistic appearance manifolds . In 2003 IEEE Computer Society Conference on Computer Vision and Pattern Recognition, 2003. Proceedings. , Vol. 1 . IEEE, I--I. Kuang-Chih Lee, Jeffrey Ho, Ming-Hsuan Yang, and David Kriegman. 2003. Video-based face recognition using probabilistic appearance manifolds. In 2003 IEEE Computer Society Conference on Computer Vision and Pattern Recognition, 2003. Proceedings., Vol. 1. IEEE, I--I."},{"key":"e_1_3_2_2_27_1","volume-title":"Asian conference on computer vision. Springer, 17--33","author":"Li Haoxiang","year":"2014","unstructured":"Haoxiang Li , Gang Hua , Xiaohui Shen , Zhe Lin , and Jonathan Brandt . 2014 . Eigen-pep for video face recognition . In Asian conference on computer vision. Springer, 17--33 . Haoxiang Li, Gang Hua, Xiaohui Shen, Zhe Lin, and Jonathan Brandt. 2014. Eigen-pep for video face recognition. In Asian conference on computer vision. Springer, 17--33."},{"key":"e_1_3_2_2_28_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR42600.2020.01332"},{"key":"e_1_3_2_2_29_1","doi-asserted-by":"publisher","DOI":"10.1109\/ICCV48922.2021.00865"},{"key":"e_1_3_2_2_30_1","volume-title":"2011 International Conference on Computer Vision. IEEE, 2486--2493","author":"Liu Lingqiao","year":"2011","unstructured":"Lingqiao Liu , Lei Wang , and Xinwang Liu . 2011 . In defense of soft-assignment coding . In 2011 International Conference on Computer Vision. IEEE, 2486--2493 . Lingqiao Liu, Lei Wang, and Xinwang Liu. 2011. In defense of soft-assignment coding. In 2011 International Conference on Computer Vision. IEEE, 2486--2493."},{"key":"e_1_3_2_2_31_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR.2017.499"},{"key":"e_1_3_2_2_32_1","doi-asserted-by":"publisher","DOI":"10.1109\/ICCVW.2019.00128"},{"key":"e_1_3_2_2_33_1","doi-asserted-by":"publisher","DOI":"10.1145\/1102351.1102420"},{"key":"e_1_3_2_2_34_1","doi-asserted-by":"publisher","DOI":"10.1109\/ICB2018.2018.00033"},{"key":"e_1_3_2_2_35_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPRW.2017.250"},{"key":"e_1_3_2_2_36_1","volume-title":"Interferences in match kernels","author":"Murray Naila","year":"2016","unstructured":"Naila Murray , Herv\u00e9 J\u00e9gou , Florent Perronnin , and Andrew Zisserman . 2016. Interferences in match kernels . IEEE transactions on pattern analysis and machine intelligence, Vol. 39 , 9 ( 2016 ), 1797--1810. Naila Murray, Herv\u00e9 J\u00e9gou, Florent Perronnin, and Andrew Zisserman. 2016. Interferences in match kernels. IEEE transactions on pattern analysis and machine intelligence, Vol. 39, 9 (2016), 1797--1810."},{"key":"e_1_3_2_2_37_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR.2014.317"},{"key":"e_1_3_2_2_38_1","doi-asserted-by":"publisher","DOI":"10.1109\/ICCV.2017.180"},{"key":"e_1_3_2_2_39_1","doi-asserted-by":"publisher","DOI":"10.1109\/ICCV.2017.408"},{"key":"e_1_3_2_2_40_1","volume-title":"International conference on machine learning. PMLR, 1431--1439","author":"Reed Scott","year":"2014","unstructured":"Scott Reed , Kihyuk Sohn , Yuting Zhang , and Honglak Lee . 2014 . Learning to disentangle factors of variation with manifold interaction . In International conference on machine learning. PMLR, 1431--1439 . Scott Reed, Kihyuk Sohn, Yuting Zhang, and Honglak Lee. 2014. Learning to disentangle factors of variation with manifold interaction. In International conference on machine learning. PMLR, 1431--1439."},{"key":"e_1_3_2_2_41_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR.2015.7298682"},{"volume-title":"Frontal to profile face verification in the wild. In 2016 IEEE winter conference on applications of computer vision (WACV)","author":"Sengupta Soumyadip","key":"e_1_3_2_2_42_1","unstructured":"Soumyadip Sengupta , Jun-Cheng Chen , Carlos Castillo , Vishal M Patel , Rama Chellappa , and David W Jacobs . 2016. Frontal to profile face verification in the wild. In 2016 IEEE winter conference on applications of computer vision (WACV) . IEEE , 1--9. Soumyadip Sengupta, Jun-Cheng Chen, Carlos Castillo, Vishal M Patel, Rama Chellappa, and David W Jacobs. 2016. Frontal to profile face verification in the wild. In 2016 IEEE winter conference on applications of computer vision (WACV). IEEE, 1--9."},{"key":"e_1_3_2_2_43_1","volume-title":"Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition. 605--613","author":"Shi Miaojing","year":"2015","unstructured":"Miaojing Shi , Yannis Avrithis , and Herv\u00e9 J\u00e9gou . 2015 . Early burst detection for memory-efficient image retrieval . In Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition. 605--613 . Miaojing Shi, Yannis Avrithis, and Herv\u00e9 J\u00e9gou. 2015. Early burst detection for memory-efficient image retrieval. In Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition. 605--613."},{"key":"e_1_3_2_2_44_1","volume-title":"IEEE","author":"Stylianou Abby","year":"2019","unstructured":"Abby Stylianou , Richard Souvenir , and Robert Pless . 2019 . Visualizing deep similarity networks. In 2019 IEEE winter conference on applications of computer vision (WACV) . IEEE , 2029--2037. Abby Stylianou, Richard Souvenir, and Robert Pless. 2019. Visualizing deep similarity networks. In 2019 IEEE winter conference on applications of computer vision (WACV). IEEE, 2029--2037."},{"key":"e_1_3_2_2_45_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR.2015.7298907"},{"key":"e_1_3_2_2_46_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR.2014.220"},{"key":"e_1_3_2_2_47_1","volume-title":"Separating style and content with bilinear models. Neural computation","author":"Tenenbaum Joshua B","year":"2000","unstructured":"Joshua B Tenenbaum and William T Freeman . 2000. Separating style and content with bilinear models. Neural computation , Vol. 12 , 6 ( 2000 ), 1247--1283. Joshua B Tenenbaum and William T Freeman. 2000. Separating style and content with bilinear models. Neural computation, Vol. 12, 6 (2000), 1247--1283."},{"key":"e_1_3_2_2_48_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR.2013.119"},{"key":"e_1_3_2_2_49_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR.2017.141"},{"key":"e_1_3_2_2_50_1","doi-asserted-by":"publisher","DOI":"10.1007\/978-3-540-88693-8_52"},{"key":"e_1_3_2_2_51_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR.2019.00364"},{"key":"e_1_3_2_2_52_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR.2018.00552"},{"key":"e_1_3_2_2_53_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR.2015.7298816"},{"key":"e_1_3_2_2_54_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR.2017.609"},{"key":"e_1_3_2_2_55_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPRW.2017.87"},{"key":"e_1_3_2_2_56_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR.2011.5995566"},{"key":"e_1_3_2_2_57_1","doi-asserted-by":"publisher","DOI":"10.1109\/ICCV48922.2021.00921"},{"key":"e_1_3_2_2_58_1","volume-title":"Inducing predictive uncertainty estimation for face recognition. arXiv preprint arXiv:2009.00603","author":"Xie Weidi","year":"2020","unstructured":"Weidi Xie , Jeffrey Byrne , and Andrew Zisserman . 2020. Inducing predictive uncertainty estimation for face recognition. arXiv preprint arXiv:2009.00603 ( 2020 ). Weidi Xie, Jeffrey Byrne, and Andrew Zisserman. 2020. Inducing predictive uncertainty estimation for face recognition. arXiv preprint arXiv:2009.00603 (2020)."},{"key":"e_1_3_2_2_59_1","doi-asserted-by":"publisher","DOI":"10.1007\/978-3-030-01252-6_48"},{"key":"e_1_3_2_2_60_1","volume-title":"Multicolumn networks for face recognition. BMVC","author":"Xie Weidi","year":"2018","unstructured":"Weidi Xie and Andrew Zisserman . 2018. Multicolumn networks for face recognition. BMVC ( 2018 ). Weidi Xie and Andrew Zisserman. 2018. Multicolumn networks for face recognition. BMVC (2018)."},{"key":"e_1_3_2_2_61_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR.2017.554"},{"key":"e_1_3_2_2_62_1","doi-asserted-by":"publisher","DOI":"10.1109\/LSP.2016.2603342"},{"key":"e_1_3_2_2_63_1","doi-asserted-by":"publisher","DOI":"10.1007\/978-3-030-58607-2_1"},{"key":"e_1_3_2_2_64_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR.2019.00484"},{"key":"e_1_3_2_2_65_1","doi-asserted-by":"publisher","DOI":"10.1609\/aaai.v33i01.33019251"},{"key":"e_1_3_2_2_66_1","first-page":"7","article-title":"Cross-pose lfw: A database for studying cross-pose face recognition in unconstrained environments. Beijing University of Posts and Telecommunications","volume":"5","author":"Zheng Tianyue","year":"2018","unstructured":"Tianyue Zheng and Weihong Deng . 2018 . Cross-pose lfw: A database for studying cross-pose face recognition in unconstrained environments. Beijing University of Posts and Telecommunications , Tech. Rep , Vol. 5 (2018), 7 . Tianyue Zheng and Weihong Deng. 2018. Cross-pose lfw: A database for studying cross-pose face recognition in unconstrained environments. Beijing University of Posts and Telecommunications, Tech. Rep, Vol. 5 (2018), 7.","journal-title":"Tech. Rep"},{"key":"e_1_3_2_2_67_1","volume-title":"Cross-age lfw: A database for studying cross-age face recognition in unconstrained environments. arXiv preprint arXiv:1708.08197","author":"Zheng Tianyue","year":"2017","unstructured":"Tianyue Zheng , Weihong Deng , and Jiani Hu. 2017. Cross-age lfw: A database for studying cross-age face recognition in unconstrained environments. arXiv preprint arXiv:1708.08197 ( 2017 ). Tianyue Zheng, Weihong Deng, and Jiani Hu. 2017. Cross-age lfw: A database for studying cross-age face recognition in unconstrained environments. arXiv preprint arXiv:1708.08197 (2017)."},{"key":"e_1_3_2_2_68_1","volume-title":"Asian conference on computer vision. Springer, 35--50","author":"Zhong Yujie","year":"2018","unstructured":"Yujie Zhong , Relja Arandjelovi\u0107 , and Andrew Zisserman . 2018 . Ghostvlad for set-based face recognition . In Asian conference on computer vision. Springer, 35--50 . Yujie Zhong, Relja Arandjelovi\u0107, and Andrew Zisserman. 2018. Ghostvlad for set-based face recognition. In Asian conference on computer vision. Springer, 35--50."}],"event":{"name":"MM '22: The 30th ACM International Conference on Multimedia","sponsor":["SIGMM ACM Special Interest Group on Multimedia"],"location":"Lisboa Portugal","acronym":"MM '22"},"container-title":["Proceedings of the 30th ACM International Conference on Multimedia"],"original-title":[],"link":[{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/3503161.3547794","content-type":"unspecified","content-version":"vor","intended-application":"text-mining"},{"URL":"https:\/\/dl.acm.org\/doi\/pdf\/10.1145\/3503161.3547794","content-type":"unspecified","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2025,6,17]],"date-time":"2025-06-17T19:02:34Z","timestamp":1750186954000},"score":1,"resource":{"primary":{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/3503161.3547794"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2022,10,10]]},"references-count":68,"alternative-id":["10.1145\/3503161.3547794","10.1145\/3503161"],"URL":"https:\/\/doi.org\/10.1145\/3503161.3547794","relation":{},"subject":[],"published":{"date-parts":[[2022,10,10]]},"assertion":[{"value":"2022-10-10","order":2,"name":"published","label":"Published","group":{"name":"publication_history","label":"Publication History"}}]}}