{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2026,7,28]],"date-time":"2026-07-28T14:46:36Z","timestamp":1785249996288,"version":"3.55.0"},"publisher-location":"New York, NY, USA","reference-count":63,"publisher":"ACM","license":[{"start":{"date-parts":[[2020,10,12]],"date-time":"2020-10-12T00:00:00Z","timestamp":1602460800000},"content-version":"vor","delay-in-days":0,"URL":"https:\/\/www.acm.org\/publications\/policies\/copyright_policy#Background"}],"funder":[{"name":"Shenzhen Municipal Development and Reform Commission","award":["No. HT-JD-CXY-201904"],"award-info":[{"award-number":["No. HT-JD-CXY-201904"]}]}],"content-domain":{"domain":["dl.acm.org"],"crossmark-restriction":true},"short-container-title":[],"published-print":{"date-parts":[[2020,10,12]]},"DOI":"10.1145\/3394171.3413529","type":"proceedings-article","created":{"date-parts":[[2020,10,12]],"date-time":"2020-10-12T13:10:18Z","timestamp":1602508218000},"page":"2463-2471","update-policy":"https:\/\/doi.org\/10.1145\/crossmark-policy","source":"Crossref","is-referenced-by-count":82,"title":["Cluster Attention Contrast for Video Anomaly Detection"],"prefix":"10.1145","author":[{"given":"Ziming","family":"Wang","sequence":"first","affiliation":[{"name":"ADSPLAB, School of ECE, Peking University, ShenZhen, China"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Yuexian","family":"Zou","sequence":"additional","affiliation":[{"name":"ADSPLAB, School of ECE, Peking University, ShenZhen, China"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Zeming","family":"Zhang","sequence":"additional","affiliation":[{"name":"Harbin institute of technology, Harbin, China"}],"role":[{"vocabulary":"crossref","role":"author"}]}],"member":"320","published-online":{"date-parts":[[2020,10,12]]},"reference":[{"key":"e_1_3_2_2_1_1","volume-title":"Latent Space Autoregression for Novelty Detection. 2019 IEEE\/CVF Conference on Computer Vision and Pattern Recognition (CVPR)","author":"Abati Davide","year":"2018","unstructured":"Davide Abati , Angelo Porrello , Simone Calderara , and Rita Cucchiara . 2018 . Latent Space Autoregression for Novelty Detection. 2019 IEEE\/CVF Conference on Computer Vision and Pattern Recognition (CVPR) (2018), 481--490. Davide Abati, Angelo Porrello, Simone Calderara, and Rita Cucchiara. 2018. Latent Space Autoregression for Novelty Detection. 2019 IEEE\/CVF Conference on Computer Vision and Pattern Recognition (CVPR) (2018), 481--490."},{"key":"e_1_3_2_2_2_1","doi-asserted-by":"publisher","DOI":"10.1109\/TPAMI.2007.70825"},{"key":"e_1_3_2_2_3_1","doi-asserted-by":"publisher","DOI":"10.1109\/ICCV.2011.6126525"},{"key":"e_1_3_2_2_4_1","doi-asserted-by":"crossref","unstructured":"Yoshua Bengio Pascal Lamblin Dan Popovici and Hugo Larochelle. 2006. Greedy Layer-Wise Training of Deep Networks. In NIPS. Yoshua Bengio Pascal Lamblin Dan Popovici and Hugo Larochelle. 2006. Greedy Layer-Wise Training of Deep Networks. In NIPS.","DOI":"10.7551\/mitpress\/7503.003.0024"},{"key":"e_1_3_2_2_5_1","volume-title":"Hinton","author":"Chen Ting","year":"2020","unstructured":"Ting Chen , Simon Kornblith , Mohammad Norouzi , and Geoffrey E . Hinton . 2020 . A Simple Framework for Contrastive Learning of Visual Representations. ArXiv abs\/2002.05709 (2020). Ting Chen, Simon Kornblith, Mohammad Norouzi, and Geoffrey E. Hinton. 2020. A Simple Framework for Contrastive Learning of Visual Representations. ArXiv abs\/2002.05709 (2020)."},{"key":"e_1_3_2_2_6_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR.2015.7298909"},{"key":"e_1_3_2_2_7_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR.2011.5995434"},{"key":"e_1_3_2_2_8_1","doi-asserted-by":"crossref","unstructured":"A. Dempster N. Laird and D. Rubin. 1977. Maximum likelihood estimation from incomplete data via the em algorithm. A. Dempster N. Laird and D. Rubin. 1977. Maximum likelihood estimation from incomplete data via the em algorithm.","DOI":"10.1111\/j.2517-6161.1977.tb01600.x"},{"key":"e_1_3_2_2_9_1","doi-asserted-by":"publisher","DOI":"10.1109\/ICCV.2017.612"},{"key":"e_1_3_2_2_10_1","doi-asserted-by":"publisher","DOI":"10.1109\/ICCV.2015.167"},{"key":"e_1_3_2_2_11_1","volume-title":"Dutta and Bonny Banerjee","author":"Jayanta","year":"2015","unstructured":"Jayanta K. Dutta and Bonny Banerjee . 2015 . Online Detection of Abnormal Events Using Incremental Coding Length. In AAAI. Jayanta K. Dutta and Bonny Banerjee. 2015. Online Detection of Abnormal Events Using Incremental Coding Length. In AAAI."},{"key":"e_1_3_2_2_12_1","doi-asserted-by":"publisher","DOI":"10.1145\/2964284.2967290"},{"key":"e_1_3_2_2_13_1","volume-title":"Learning Temporal Regularity in Video Sequences. 2016 IEEE Conference on Computer Vision and Pattern Recognition (CVPR)","author":"Hasan Mahmudul","year":"2016","unstructured":"Mahmudul Hasan , Jonghyun Choi , Jan Neumann , Amit K. Roy-Chowdhury , and Larry S. Davis . 2016 . Learning Temporal Regularity in Video Sequences. 2016 IEEE Conference on Computer Vision and Pattern Recognition (CVPR) ( 2016 ), 733--742. Mahmudul Hasan, Jonghyun Choi, Jan Neumann, Amit K. Roy-Chowdhury, and Larry S. Davis. 2016. Learning Temporal Regularity in Video Sequences. 2016 IEEE Conference on Computer Vision and Pattern Recognition (CVPR) (2016), 733--742."},{"key":"e_1_3_2_2_14_1","volume-title":"Girshick","author":"He Kaiming","year":"2019","unstructured":"Kaiming He , Haoqi Fan , Yuxin Wu , Saining Xie , and Ross B . Girshick . 2019 . Momentum Contrast for Unsupervised Visual Representation Learning. ArXiv abs\/1911.05722 (2019). Kaiming He, Haoqi Fan, Yuxin Wu, Saining Xie, and Ross B. Girshick. 2019. Momentum Contrast for Unsupervised Visual Representation Learning. ArXiv abs\/1911.05722 (2019)."},{"key":"e_1_3_2_2_15_1","volume-title":"Hinton and Ruslan Salakhutdinov","author":"Geoffrey","year":"2006","unstructured":"Geoffrey E. Hinton and Ruslan Salakhutdinov . 2006 . Reducing the dimensionality of data with neural networks. Science 313 5786 (2006), 504--7. Geoffrey E. Hinton and Ruslan Salakhutdinov. 2006. Reducing the dimensionality of data with neural networks. Science 313 5786 (2006), 504--7."},{"key":"e_1_3_2_2_16_1","doi-asserted-by":"publisher","DOI":"10.1109\/ICCV.2009.5459342"},{"key":"e_1_3_2_2_17_1","doi-asserted-by":"publisher","DOI":"10.1109\/TMM.2017.2745702"},{"key":"e_1_3_2_2_18_1","volume-title":"Deep Multimodal Clustering for Unsupervised Audiovisual Learning. 2019 IEEE\/CVF Conference on Computer Vision and Pattern Recognition (CVPR)","author":"Hu Di","year":"2018","unstructured":"Di Hu , Feiping Nie , and Xuelong Li . 2018 . Deep Multimodal Clustering for Unsupervised Audiovisual Learning. 2019 IEEE\/CVF Conference on Computer Vision and Pattern Recognition (CVPR) (2018), 9240--9249. Di Hu, Feiping Nie, and Xuelong Li. 2018. Deep Multimodal Clustering for Unsupervised Audiovisual Learning. 2019 IEEE\/CVF Conference on Computer Vision and Pattern Recognition (CVPR) (2018), 9240--9249."},{"key":"e_1_3_2_2_19_1","volume-title":"Learning Discrete Representations via Information Maximizing Self-Augmented Training. ArXiv abs\/1702.08720","author":"Hu Weihua","year":"2017","unstructured":"Weihua Hu , Takeru Miyato , Seiya Tokui , Eiichi Matsumoto , and Masashi Sugiyama . 2017. Learning Discrete Representations via Information Maximizing Self-Augmented Training. ArXiv abs\/1702.08720 ( 2017 ). Weihua Hu, Takeru Miyato, Seiya Tokui, Eiichi Matsumoto, and Masashi Sugiyama. 2017. Learning Discrete Representations via Information Maximizing Self-Augmented Training. ArXiv abs\/1702.08720 (2017)."},{"key":"e_1_3_2_2_20_1","volume-title":"Batch Normalization: Accelerating Deep Network Training by Reducing Internal Covariate Shift. ArXiv abs\/1502.03167","author":"Ioffe Sergey","year":"2015","unstructured":"Sergey Ioffe and Christian Szegedy . 2015 . Batch Normalization: Accelerating Deep Network Training by Reducing Internal Covariate Shift. ArXiv abs\/1502.03167 (2015). Sergey Ioffe and Christian Szegedy. 2015. Batch Normalization: Accelerating Deep Network Training by Reducing Internal Covariate Shift. ArXiv abs\/1502.03167 (2015)."},{"key":"e_1_3_2_2_21_1","volume-title":"Object-Centric Auto-Encoders and Dummy Anomalies for Abnormal Event Detection in Video. 2019 IEEE\/CVF Conference on Computer Vision and Pattern Recognition (CVPR)","author":"Ionescu Radu Tudor","year":"2019","unstructured":"Radu Tudor Ionescu , Fahad Shahbaz Khan , Mariana-Iuliana Georgescu , and Ling Shao . 2019 . Object-Centric Auto-Encoders and Dummy Anomalies for Abnormal Event Detection in Video. 2019 IEEE\/CVF Conference on Computer Vision and Pattern Recognition (CVPR) (2019), 7834--7843. Radu Tudor Ionescu, Fahad Shahbaz Khan, Mariana-Iuliana Georgescu, and Ling Shao. 2019. Object-Centric Auto-Encoders and Dummy Anomalies for Abnormal Event Detection in Video. 2019 IEEE\/CVF Conference on Computer Vision and Pattern Recognition (CVPR) (2019), 7834--7843."},{"key":"e_1_3_2_2_22_1","volume-title":"Detecting Abnormal Events in Video Using Narrowed Normality Clusters. 2019 IEEE Winter Conference on Applications of Computer Vision (WACV) (2019)","author":"Ionescu Radu Tudor","year":"2019","unstructured":"Radu Tudor Ionescu , Sorina Smeureanu , Marius Popescu , and Bogdan Alexe . 2019 . Detecting Abnormal Events in Video Using Narrowed Normality Clusters. 2019 IEEE Winter Conference on Applications of Computer Vision (WACV) (2019) , 1951--1960. Radu Tudor Ionescu, Sorina Smeureanu, Marius Popescu, and Bogdan Alexe. 2019. Detecting Abnormal Events in Video Using Narrowed Normality Clusters. 2019 IEEE Winter Conference on Applications of Computer Vision (WACV) (2019), 1951--1960."},{"key":"e_1_3_2_2_23_1","doi-asserted-by":"publisher","DOI":"10.1007\/s11263-017-1001-2"},{"key":"e_1_3_2_2_24_1","volume-title":"Reid","author":"Ji Pan","year":"2017","unstructured":"Pan Ji , Tong Zhang , Hongdong Li , Mathieu Salzmann , and Ian D . Reid . 2017 . Deep Subspace Clustering Networks. In NIPS. Pan Ji, Tong Zhang, Hongdong Li, Mathieu Salzmann, and Ian D. Reid. 2017. Deep Subspace Clustering Networks. In NIPS."},{"key":"e_1_3_2_2_25_1","volume-title":"Invariant Information Clustering for Unsupervised Image Classification and Segmentation. 2019 IEEE\/CVF International Conference on Computer Vision (ICCV)","author":"Ji Xu","year":"2018","unstructured":"Xu Ji , Andrea Vedaldi , and Jo\u00e3o F. Henriques . 2018 . Invariant Information Clustering for Unsupervised Image Classification and Segmentation. 2019 IEEE\/CVF International Conference on Computer Vision (ICCV) ( 2018 ), 9864--9873. Xu Ji, Andrea Vedaldi, and Jo\u00e3o F. Henriques. 2018. Invariant Information Clustering for Unsupervised Image Classification and Segmentation. 2019 IEEE\/CVF International Conference on Computer Vision (ICCV) (2018), 9864--9873."},{"key":"e_1_3_2_2_26_1","volume-title":"The Kinetics Human Action Video Dataset. ArXiv abs\/1705.06950","author":"Kay Will","year":"2017","unstructured":"Will Kay , Jo\u00e3o Carreira , Karen Simonyan , Brian Zhang , Chloe Hillier , Sudheendra Vijayanarasimhan , Fabio Viola , Tim Green , Trevor Back , Apostol Natsev , Mustafa Suleyman , and Andrew Zisserman . 2017. The Kinetics Human Action Video Dataset. ArXiv abs\/1705.06950 ( 2017 ). Will Kay, Jo\u00e3o Carreira, Karen Simonyan, Brian Zhang, Chloe Hillier, Sudheendra Vijayanarasimhan, Fabio Viola, Tim Green, Trevor Back, Apostol Natsev, Mustafa Suleyman, and Andrew Zisserman. 2017. The Kinetics Human Action Video Dataset. ArXiv abs\/1705.06950 (2017)."},{"key":"e_1_3_2_2_27_1","doi-asserted-by":"crossref","unstructured":"Jaechul Kim and Kristen Grauman. 2009. Observe locally infer globally: A spacetime MRF for detecting abnormal activities with incremental updates. In CVPR. Jaechul Kim and Kristen Grauman. 2009. Observe locally infer globally: A spacetime MRF for detecting abnormal activities with incremental updates. In CVPR.","DOI":"10.1109\/CVPRW.2009.5206569"},{"key":"e_1_3_2_2_28_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR.2009.5206771"},{"key":"e_1_3_2_2_29_1","article-title":"Anomaly Detection in Video Surveillance via Gaussian","volume":"29","author":"Li Nannan","year":"2015","unstructured":"Nannan Li , Xinyu Wu , Huiwen Guo , Dan Xu , Yongsheng Ou , and Yen-Lun Chen . 2015 . Anomaly Detection in Video Surveillance via Gaussian Process. Int. J. Pattern Recognit. Artif. Intell. 29 (2015), 1555011:1--1555011:25. Nannan Li, Xinyu Wu, Huiwen Guo, Dan Xu, Yongsheng Ou, and Yen-Lun Chen. 2015. Anomaly Detection in Video Surveillance via Gaussian Process. Int. J. Pattern Recognit. Artif. Intell. 29 (2015), 1555011:1--1555011:25.","journal-title":"Process. Int. J. Pattern Recognit. Artif. Intell."},{"key":"e_1_3_2_2_30_1","doi-asserted-by":"publisher","DOI":"10.1109\/TPAMI.2013.111"},{"key":"e_1_3_2_2_31_1","volume-title":"Berg","author":"Liu Wei","year":"2016","unstructured":"Wei Liu , Dragomir Anguelov , Dumitru Erhan , Christian Szegedy , Scott E. Reed , Cheng-Yang Fu , and Alexander C . Berg . 2016 . SSD : Single Shot MultiBox Detector. In ECCV. Wei Liu, Dragomir Anguelov, Dumitru Erhan, Christian Szegedy, Scott E. Reed, Cheng-Yang Fu, and Alexander C. Berg. 2016. SSD: Single Shot MultiBox Detector. In ECCV."},{"key":"e_1_3_2_2_32_1","volume-title":"Future Frame Prediction for Anomaly Detection - A New Baseline. 2018 IEEE\/CVF Conference on Computer Vision and Pattern Recognition","author":"Liu Wen","year":"2018","unstructured":"Wen Liu , Weixin Luo , Dongze Lian , and Shenghua Gao . 2018 . Future Frame Prediction for Anomaly Detection - A New Baseline. 2018 IEEE\/CVF Conference on Computer Vision and Pattern Recognition (2018), 6536--6545. Wen Liu, Weixin Luo, Dongze Lian, and Shenghua Gao. 2018. Future Frame Prediction for Anomaly Detection - A New Baseline. 2018 IEEE\/CVF Conference on Computer Vision and Pattern Recognition (2018), 6536--6545."},{"key":"e_1_3_2_2_33_1","unstructured":"Yusha Liu Chun-Liang Li and Barnab\u00e1s P\u00f3czos. 2018. Classifier Two Sample Test for Video Anomaly Detections. In BMVC. Yusha Liu Chun-Liang Li and Barnab\u00e1s P\u00f3czos. 2018. Classifier Two Sample Test for Video Anomaly Detections. In BMVC."},{"key":"e_1_3_2_2_34_1","doi-asserted-by":"publisher","DOI":"10.1109\/ICCV.2013.338"},{"key":"e_1_3_2_2_35_1","doi-asserted-by":"publisher","DOI":"10.1109\/ICCV.2017.45"},{"key":"e_1_3_2_2_36_1","doi-asserted-by":"crossref","unstructured":"Jonathan Masci Ueli Meier Dan C. Ciresan and J\u00fcrgen Schmidhuber. 2011. Stacked Convolutional Auto-Encoders for Hierarchical Feature Extraction. In ICANN. Jonathan Masci Ueli Meier Dan C. Ciresan and J\u00fcrgen Schmidhuber. 2011. Stacked Convolutional Auto-Encoders for Hierarchical Feature Extraction. In ICANN.","DOI":"10.1007\/978-3-642-21735-7_7"},{"key":"e_1_3_2_2_37_1","volume-title":"Savakis","author":"Medel Jefferson Ryan","year":"2016","unstructured":"Jefferson Ryan Medel and Andreas E . Savakis . 2016 . Anomaly Detection in Video Using Predictive Convolutional Long Short-Term Memory Networks. ArXiv abs\/1612.00390 (2016). Jefferson Ryan Medel and Andreas E. Savakis. 2016. Anomaly Detection in Video Using Predictive Convolutional Long Short-Term Memory Networks. ArXiv abs\/1612.00390 (2016)."},{"key":"e_1_3_2_2_38_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR.2009.5206641"},{"key":"e_1_3_2_2_39_1","doi-asserted-by":"crossref","unstructured":"Mehdi Noroozi and Paolo Favaro. 2016. Unsupervised Learning of Visual Representations by Solving Jigsaw Puzzles. In ECCV. Mehdi Noroozi and Paolo Favaro. 2016. Unsupervised Learning of Visual Representations by Solving Jigsaw Puzzles. In ECCV.","DOI":"10.1007\/978-3-319-46466-4_5"},{"key":"e_1_3_2_2_40_1","doi-asserted-by":"publisher","DOI":"10.1109\/ICCV.2017.628"},{"key":"e_1_3_2_2_41_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR.2017.638"},{"key":"e_1_3_2_2_42_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR.2016.278"},{"key":"e_1_3_2_2_43_1","volume-title":"Moeslund","author":"Ren Huamin","year":"2015","unstructured":"Huamin Ren , Weifeng Liu , S\u00f8ren I. Olsen , Sergio Escalera , and Thomas B . Moeslund . 2015 . Unsupervised Behavior-Specific Dictionary Learning for Abnormal Event Detection. In BMVC. Huamin Ren, Weifeng Liu, S\u00f8ren I. Olsen, Sergio Escalera, and Thomas B. Moeslund. 2015. Unsupervised Behavior-Specific Dictionary Learning for Abnormal Event Detection. In BMVC."},{"key":"e_1_3_2_2_44_1","doi-asserted-by":"publisher","DOI":"10.1214\/aoms\/1177729586"},{"key":"e_1_3_2_2_45_1","doi-asserted-by":"publisher","DOI":"10.1162\/089976601750264965"},{"key":"e_1_3_2_2_46_1","volume-title":"Contrastive Multiview Coding. ArXiv abs\/1906.05849","author":"Tian Yonglong","year":"2019","unstructured":"Yonglong Tian , Dilip Krishnan , and Phillip Isola . 2019. Contrastive Multiview Coding. ArXiv abs\/1906.05849 ( 2019 ). Yonglong Tian, Dilip Krishnan, and Phillip Isola. 2019. Contrastive Multiview Coding. ArXiv abs\/1906.05849 (2019)."},{"key":"e_1_3_2_2_47_1","volume-title":"Hogg","author":"Tran Hanh","year":"2017","unstructured":"Hanh Tran and David C . Hogg . 2017 . Anomaly Detection using a Convolutional Winner-Take-All Autoencoder. In BMVC. Hanh Tran and David C. Hogg. 2017. Anomaly Detection using a Convolutional Winner-Take-All Autoencoder. In BMVC."},{"key":"e_1_3_2_2_48_1","volume-title":"Representation Learning with Contrastive Predictive Coding. ArXiv abs\/1807.03748","author":"van den Oord A\u00e4ron","year":"2018","unstructured":"A\u00e4ron van den Oord , Yazhe Li , and Oriol Vinyals . 2018. Representation Learning with Contrastive Predictive Coding. ArXiv abs\/1807.03748 ( 2018 ). A\u00e4ron van den Oord, Yazhe Li, and Oriol Vinyals. 2018. Representation Learning with Contrastive Predictive Coding. ArXiv abs\/1807.03748 (2018)."},{"key":"e_1_3_2_2_49_1","doi-asserted-by":"publisher","DOI":"10.5555\/1756006.1953039"},{"key":"e_1_3_2_2_50_1","volume-title":"Unsupervised Learning of Visual Representations Using Videos. 2015 IEEE International Conference on Computer Vision (ICCV)","author":"Wang Xiaolong","year":"2015","unstructured":"Xiaolong Wang and Abhinav Gupta . 2015 . Unsupervised Learning of Visual Representations Using Videos. 2015 IEEE International Conference on Computer Vision (ICCV) (2015), 2794--2802. Xiaolong Wang and Abhinav Gupta. 2015. Unsupervised Learning of Visual Representations Using Videos. 2015 IEEE International Conference on Computer Vision (ICCV) (2015), 2794--2802."},{"key":"e_1_3_2_2_51_1","volume-title":"Transitive Invariance for Self-Supervised Visual Representation Learning. 2017 IEEE International Conference on Computer Vision (ICCV)","author":"Wang Xiaolong","year":"2017","unstructured":"Xiaolong Wang , Kaiming He , and Abhinav Gupta . 2017 . Transitive Invariance for Self-Supervised Visual Representation Learning. 2017 IEEE International Conference on Computer Vision (ICCV) (2017), 1338--1347. Xiaolong Wang, Kaiming He, and Abhinav Gupta. 2017. Transitive Invariance for Self-Supervised Visual Representation Learning. 2017 IEEE International Conference on Computer Vision (ICCV) (2017), 1338--1347."},{"key":"e_1_3_2_2_52_1","unstructured":"Junyuan Xie Ross B. Girshick and Ali Farhadi. 2016. Unsupervised Deep Embedding for Clustering Analysis. In ICML. Junyuan Xie Ross B. Girshick and Ali Farhadi. 2016. Unsupervised Deep Embedding for Clustering Analysis. In ICML."},{"key":"e_1_3_2_2_53_1","doi-asserted-by":"publisher","DOI":"10.1016\/j.cviu.2016.10.010"},{"key":"e_1_3_2_2_54_1","volume-title":"Joint Unsupervised Learning of Deep Representations and Image Clusters. 2016 IEEE Conference on Computer Vision and Pattern Recognition (CVPR)","author":"Yang Jianwei","year":"2016","unstructured":"Jianwei Yang , Devi Parikh , and Dhruv Batra . 2016 . Joint Unsupervised Learning of Deep Representations and Image Clusters. 2016 IEEE Conference on Computer Vision and Pattern Recognition (CVPR) (2016), 5147--5156. Jianwei Yang, Devi Parikh, and Dhruv Batra. 2016. Joint Unsupervised Learning of Deep Representations and Image Clusters. 2016 IEEE Conference on Computer Vision and Pattern Recognition (CVPR) (2016), 5147--5156."},{"key":"e_1_3_2_2_55_1","doi-asserted-by":"publisher","DOI":"10.1145\/3343031.3350899"},{"key":"e_1_3_2_2_56_1","doi-asserted-by":"publisher","DOI":"10.1145\/3343031.3350876"},{"key":"e_1_3_2_2_57_1","volume-title":"Self-Supervised Convolutional Subspace Clustering Network. 2019 IEEE\/CVF Conference on Computer Vision and Pattern Recognition (CVPR)","author":"Zhang Junjian","year":"2019","unstructured":"Junjian Zhang , Chun-Guang Li , Chong You , Xianbiao Qi , Honggang Zhang , Jun Guo , and Zhouchen Lin . 2019 . Self-Supervised Convolutional Subspace Clustering Network. 2019 IEEE\/CVF Conference on Computer Vision and Pattern Recognition (CVPR) (2019), 5468--5477. Junjian Zhang, Chun-Guang Li, Chong You, Xianbiao Qi, Honggang Zhang, Jun Guo, and Zhouchen Lin. 2019. Self-Supervised Convolutional Subspace Clustering Network. 2019 IEEE\/CVF Conference on Computer Vision and Pattern Recognition (CVPR) (2019), 5468--5477."},{"key":"e_1_3_2_2_58_1","volume-title":"Efros","author":"Zhang Richard","year":"2016","unstructured":"Richard Zhang , Phillip Isola , and Alexei A . Efros . 2016 . Colorful Image Colorization. ArXiv abs\/1603.08511 (2016). Richard Zhang, Phillip Isola, and Alexei A. Efros. 2016. Colorful Image Colorization. ArXiv abs\/1603.08511 (2016)."},{"key":"e_1_3_2_2_59_1","volume-title":"Reid","author":"Zhang Tong","year":"2018","unstructured":"Tong Zhang , Pan Ji , Mehrtash Harandi , Richard I. Hartley , and Ian D . Reid . 2018 . Scalable Deep k-Subspace Clustering. ArXiv abs\/1811.01045 (2018). Tong Zhang, Pan Ji, Mehrtash Harandi, Richard I. Hartley, and Ian D. Reid. 2018. Scalable Deep k-Subspace Clustering. ArXiv abs\/1811.01045 (2018)."},{"key":"e_1_3_2_2_60_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR.2011.5995524"},{"key":"e_1_3_2_2_61_1","doi-asserted-by":"publisher","DOI":"10.1145\/3123266.3123451"},{"key":"e_1_3_2_2_62_1","volume-title":"Deep Adversarial Subspace Clustering. 2018 IEEE\/CVF Conference on Computer Vision and Pattern Recognition","author":"Zhou Pan","year":"2018","unstructured":"Pan Zhou , Yunqing Hou , and Jiashi Feng . 2018 . Deep Adversarial Subspace Clustering. 2018 IEEE\/CVF Conference on Computer Vision and Pattern Recognition (2018), 1596--1604. Pan Zhou, Yunqing Hou, and Jiashi Feng. 2018. Deep Adversarial Subspace Clustering. 2018 IEEE\/CVF Conference on Computer Vision and Pattern Recognition (2018), 1596--1604."},{"key":"e_1_3_2_2_63_1","volume-title":"2017 IEEE Conference on Computer Vision and Pattern Recognition (CVPR)","author":"Zhou Tinghui","year":"2017","unstructured":"Tinghui Zhou , Matthew Brown , Noah Snavely , and David G. Lowe . 2017. Unsupervised Learning of Depth and Ego-Motion from Video . 2017 IEEE Conference on Computer Vision and Pattern Recognition (CVPR) ( 2017 ), 6612--661 Tinghui Zhou, Matthew Brown, Noah Snavely, and David G. Lowe. 2017. Unsupervised Learning of Depth and Ego-Motion from Video. 2017 IEEE Conference on Computer Vision and Pattern Recognition (CVPR) (2017), 6612--661"}],"event":{"name":"MM '20: The 28th ACM International Conference on Multimedia","location":"Seattle WA USA","acronym":"MM '20","sponsor":["SIGMM ACM Special Interest Group on Multimedia"]},"container-title":["Proceedings of the 28th ACM International Conference on Multimedia"],"original-title":[],"link":[{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/3394171.3413529","content-type":"unspecified","content-version":"vor","intended-application":"text-mining"},{"URL":"https:\/\/dl.acm.org\/doi\/pdf\/10.1145\/3394171.3413529","content-type":"unspecified","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2025,6,17]],"date-time":"2025-06-17T20:47:13Z","timestamp":1750193233000},"score":1,"resource":{"primary":{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/3394171.3413529"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2020,10,12]]},"references-count":63,"alternative-id":["10.1145\/3394171.3413529","10.1145\/3394171"],"URL":"https:\/\/doi.org\/10.1145\/3394171.3413529","relation":{},"subject":[],"published":{"date-parts":[[2020,10,12]]},"assertion":[{"value":"2020-10-12","order":2,"name":"published","label":"Published","group":{"name":"publication_history","label":"Publication History"}}]}}