{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2026,7,16]],"date-time":"2026-07-16T21:19:54Z","timestamp":1784236794101,"version":"3.55.0"},"reference-count":57,"publisher":"MDPI AG","issue":"9","license":[{"start":{"date-parts":[[2021,4,29]],"date-time":"2021-04-29T00:00:00Z","timestamp":1619654400000},"content-version":"vor","delay-in-days":0,"URL":"https:\/\/creativecommons.org\/licenses\/by\/4.0\/"}],"funder":[{"DOI":"10.13039\/501100013285","name":"Program for Professor of Special Appointment (Eastern Scholar) at Shanghai Institutions of Higher Learning","doi-asserted-by":"publisher","award":["ES2014 XX"],"award-info":[{"award-number":["ES2014 XX"]}],"id":[{"id":"10.13039\/501100013285","id-type":"DOI","asserted-by":"publisher"}]},{"name":"JSPS KAKENHI","award":["15K00159"],"award-info":[{"award-number":["15K00159"]}]}],"content-domain":{"domain":[],"crossmark-restriction":false},"short-container-title":["Sensors"],"abstract":"<jats:p>Recently, with the popularization of camera tools such as mobile phones and the rise of various short video platforms, a lot of videos are being uploaded to the Internet at all times, for which a video retrieval system with fast retrieval speed and high precision is very necessary. Therefore, content-based video retrieval (CBVR) has aroused the interest of many researchers. A typical CBVR system mainly contains the following two essential parts: video feature extraction and similarity comparison. Feature extraction of video is very challenging, previous video retrieval methods are mostly based on extracting features from single video frames, while resulting the loss of temporal information in the videos. Hashing methods are extensively used in multimedia information retrieval due to its retrieval efficiency, but most of them are currently only applied to image retrieval. In order to solve these problems in video retrieval, we build an end-to-end framework called deep supervised video hashing (DSVH), which employs a 3D convolutional neural network (CNN) to obtain spatial-temporal features of videos, then train a set of hash functions by supervised hashing to transfer the video features into binary space and get the compact binary codes of videos. Finally, we use triplet loss for network training. We conduct a lot of experiments on three public video datasets UCF-101, JHMDB and HMDB-51, and the results show that the proposed method has advantages over many state-of-the-art video retrieval methods. Compared with the DVH method, the mAP value of UCF-101 dataset is improved by 9.3%, and the minimum improvement on JHMDB dataset is also increased by 0.3%. At the same time, we also demonstrate the stability of the algorithm in the HMDB-51 dataset.<\/jats:p>","DOI":"10.3390\/s21093094","type":"journal-article","created":{"date-parts":[[2021,4,29]],"date-time":"2021-04-29T04:30:11Z","timestamp":1619670611000},"page":"3094","update-policy":"https:\/\/doi.org\/10.3390\/mdpi_crossmark_policy","source":"Crossref","is-referenced-by-count":29,"title":["A Supervised Video Hashing Method Based on a Deep 3D Convolutional Neural Network for Large-Scale Video Retrieval"],"prefix":"10.3390","volume":"21","author":[{"ORCID":"https:\/\/orcid.org\/0000-0003-3305-4489","authenticated-orcid":false,"given":"Hanqing","family":"Chen","sequence":"first","affiliation":[{"name":"Shanghai Engineering Research Center of Assistive Devices, School of Medical Instrument and Food Engineering, University of Shanghai for Science and Technology, Shanghai 200093, China"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Chunyan","family":"Hu","sequence":"additional","affiliation":[{"name":"School of Optical-Electrical and Computer Engineering, University of Shanghai for Science and Technology, Shanghai 200093, China"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Feifei","family":"Lee","sequence":"additional","affiliation":[{"name":"Shanghai Engineering Research Center of Assistive Devices, School of Medical Instrument and Food Engineering, University of Shanghai for Science and Technology, Shanghai 200093, China"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Chaowei","family":"Lin","sequence":"additional","affiliation":[{"name":"Shanghai Engineering Research Center of Assistive Devices, School of Medical Instrument and Food Engineering, University of Shanghai for Science and Technology, Shanghai 200093, China"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Wei","family":"Yao","sequence":"additional","affiliation":[{"name":"Shanghai Engineering Research Center of Assistive Devices, School of Medical Instrument and Food Engineering, University of Shanghai for Science and Technology, Shanghai 200093, China"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Lu","family":"Chen","sequence":"additional","affiliation":[{"name":"Shanghai Engineering Research Center of Assistive Devices, School of Medical Instrument and Food Engineering, University of Shanghai for Science and Technology, Shanghai 200093, China"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"ORCID":"https:\/\/orcid.org\/0000-0003-3079-9207","authenticated-orcid":false,"given":"Qiu","family":"Chen","sequence":"additional","affiliation":[{"name":"Major of Electrical Engineering and Electronics, Graduate School of Engineering, Kogakuin University, Tokyo 163-8677, Japan"}],"role":[{"vocabulary":"crossref","role":"author"}]}],"member":"1968","published-online":{"date-parts":[[2021,4,29]]},"reference":[{"key":"ref_1","doi-asserted-by":"crossref","first-page":"257","DOI":"10.1109\/TMM.2010.2046265","article-title":"An image-based approach to video copy detection with spatio-temporal post-filtering","volume":"12","author":"Douze","year":"2010","journal-title":"IEEE Trans. Multimed."},{"key":"ref_2","doi-asserted-by":"crossref","first-page":"99","DOI":"10.1007\/s13740-016-0060-9","article-title":"Content-based video recommendation system based on stylistic visual features","volume":"5","author":"Deldjoo","year":"2016","journal-title":"J. Data Semant."},{"key":"ref_3","doi-asserted-by":"crossref","unstructured":"Dong, Y., and Li, J. (2018, January 28). Video Retrieval based on Deep Convolutional Neural Network. Proceedings of the 3rd International Conference on Multimedia Systems and Signal Processing (ICMSSP), Shenzhen, China.","DOI":"10.1145\/3220162.3220168"},{"key":"ref_4","doi-asserted-by":"crossref","unstructured":"Muranoi, R., Zhao, J., Hayasaka, R., Ito, M., and Matsushita, Y. (1998, January 5). Video retrieval method using shotID for copyright protection systems. Proceedings of the Multimedia Storage and Archiving Systems III, Boston, MA, USA.","DOI":"10.1117\/12.325818"},{"key":"ref_5","doi-asserted-by":"crossref","unstructured":"Shang, L., Yang, L., Wang, F., Chan, K.P., and Hua, X.S. (2010, January 25\u201329). Real-Time Large Scale Near-Duplicate Web Video Retrieval. Proceedings of the 18th ACM International Conference on Multimedia (ACM MM), Firenze, Italy.","DOI":"10.1145\/1873951.1874021"},{"key":"ref_6","doi-asserted-by":"crossref","unstructured":"Wu, X., Hauptmann, A.G., and Ngo, C.W. (2007, January 24\u201329). Practical Elimination of Near-Duplicates from Web Video Search. Proceedings of the 15th ACM International Conference on Multimedia (ACM MM), Augsburg, Bavaria, Germany.","DOI":"10.1145\/1291233.1291280"},{"key":"ref_7","doi-asserted-by":"crossref","first-page":"91","DOI":"10.1023\/B:VISI.0000029664.99615.94","article-title":"Distinctive image features from scale-invariant keypoints","volume":"60","author":"Lowe","year":"2004","journal-title":"Int. J. Comput. Vis."},{"key":"ref_8","unstructured":"Krizhevsky, A., Sutskever, I., and Hinton, G.E. (2012, January 3\u20136). ImageNet Classification with Deep Convolutional Neural Networks. Proceedings of the Advances in Neural Information Processing Systems (NIPS), Lake Tahoe, NV, USA."},{"key":"ref_9","doi-asserted-by":"crossref","first-page":"1182","DOI":"10.1109\/TMM.2019.2942478","article-title":"Hierarchical coding of convolutional features for scene recognition","volume":"22","author":"Xie","year":"2020","journal-title":"IEEE Trans. Multimed."},{"key":"ref_10","first-page":"505","article-title":"Advanced feature fusion algorithm based on multiple convolutional neural network for scene recognition","volume":"122","author":"Chen","year":"2020","journal-title":"Comput. Model. Eng. Sci."},{"key":"ref_11","unstructured":"Ren, S., He, K., Girshick, R., and Sun, J. (2015, January 11\u201312). Faster r-Cnn: Towards Real-Time Object Detection with Region Proposal Networks. Proceedings of the Advances in Neural Information Processing Systems (NIPS), Montreal, Canada."},{"key":"ref_12","doi-asserted-by":"crossref","unstructured":"Girshick, R., Donahue, J., Darrell, T., and Malik, J. (2014, January 23\u201328). Rich Feature Hierarchies for Accurate Object Detection and Semantic Segmentation. Proceedings of the 2014 IEEE Conference on Computer Vision and Pattern Recognition (CVPR), Columbus, OH, USA.","DOI":"10.1109\/CVPR.2014.81"},{"key":"ref_13","doi-asserted-by":"crossref","unstructured":"Deng, J., Guo, J., Xue, N., and Zafeiriou, S. (2019, January 16\u201320). Arcface: Additive Angular Margin Loss for Deep Face Recognition. Proceedings of the 2019 IEEE Conference on Computer Vision and Pattern Recognition (CVPR), Long Beach, CA, USA.","DOI":"10.1109\/CVPR.2019.00482"},{"key":"ref_14","doi-asserted-by":"crossref","unstructured":"Xia, R., Pan, Y., Lai, H., Liu, C., and Yan, S. (2014, January 27\u201331). Supervised Hashing for Image Retrieval via Image Representation Learning. Proceedings of the 28th AAAI Conference on Artificial Intelligence, Quebec City, QC, Canada.","DOI":"10.1609\/aaai.v28i1.8952"},{"key":"ref_15","unstructured":"Zhao, F., Huang, Y., Wang, L., and Tan, T. (2015, January 7\u201312). Deep Semantic Ranking based Hashing for Multi-Label Image Retrieval. Proceedings of the 2015 IEEE Conference on Computer VISION and Pattern Recognition (CVPR), Boston, MA, USA."},{"key":"ref_16","first-page":"593","article-title":"IDSH: An improved deep supervised hashing method for image retrieval","volume":"121","author":"Lu","year":"2019","journal-title":"Comput. Model. Eng. Sci."},{"key":"ref_17","doi-asserted-by":"crossref","unstructured":"Tan, H.K., Ngo, C.W., Hong, R., and Chua, T.S. (2009, January 19\u201324). Scalable Detection of Partial Near-Duplicate Videos by Visual-Temporal Consistency. Proceedings of the 17th ACM International Conference on Multimedia (ACM MM), Beijing, China.","DOI":"10.1145\/1631272.1631295"},{"key":"ref_18","doi-asserted-by":"crossref","unstructured":"Douze, M., J\u00e9gou, H., Schmid, C., and P\u00e9rez, P. (2010, January 5\u201311). Compact Video Description for Copy Detection with Precise Temporal Alignment. Proceedings of the European Conference on Computer Vision (ECCV), Crete, Greece.","DOI":"10.1007\/978-3-642-15549-9_38"},{"key":"ref_19","doi-asserted-by":"crossref","unstructured":"Baraldi, L., Douze, M., Cucchiara, R., and J\u00e9gou, H. (2018, January 18\u201323). LAMV: Learning to align and match videos with kernelized temporal layers. Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition (CVPR), Salt Lake City, UT, USA.","DOI":"10.1109\/CVPR.2018.00814"},{"key":"ref_20","unstructured":"Kordopatis-Zilos, G., Papadopoulos, S., Patras, I., and Kompatsiaris, I. (November, January 27). Visil: Fine-Grained Spatio-Temporal Video Similarity Learning. Proceedings of the IEEE International Conference on Computer Vision (ICCV), Seoul, Korea."},{"key":"ref_21","doi-asserted-by":"crossref","first-page":"21","DOI":"10.1016\/j.jvcir.2018.05.013","article-title":"Learning spatial-temporal features for video copy detection by the combination of CNN and RNN","volume":"55","author":"Hu","year":"2018","journal-title":"J. Vis. Commun. Image Represent."},{"key":"ref_22","doi-asserted-by":"crossref","unstructured":"Lin, K., Yang, H.F., Hsiao, J.H., and Chen, C.S. (2015, January 7\u201312). Deep Learning of Binary Hash Codes for Fast Image Retrieval. Proceedings of the 2015 IEEE Conference on Computer Vision and Pattern Recognition Workshops (CVPRW), Boston, MA, USA.","DOI":"10.1109\/CVPRW.2015.7301269"},{"key":"ref_23","doi-asserted-by":"crossref","unstructured":"Lai, H., Pan, Y., Liu, Y., and Yan, S. (2015, January 7\u201312). Simultaneous Feature Learning and Hash Coding with Deep Neural Networks. Proceedings of the 2015 IEEE Conference on Computer Vision and Pattern Recognition (CVPR), Boston, MA, USA.","DOI":"10.1109\/CVPR.2015.7298947"},{"key":"ref_24","doi-asserted-by":"crossref","first-page":"260","DOI":"10.1016\/j.jvcir.2019.03.024","article-title":"Similarity-preserving hashing based on deep neural networks for large-scale image retrieval","volume":"61","author":"Wang","year":"2019","journal-title":"J. Vis. Commun. Image Represent."},{"key":"ref_25","doi-asserted-by":"crossref","first-page":"3330","DOI":"10.1109\/TCYB.2019.2894498","article-title":"Enhancing sketch-based image retrieval by cnn semantic re-ranking","volume":"50","author":"Wang","year":"2019","journal-title":"IEEE Trans. Cybern."},{"key":"ref_26","unstructured":"Song, J., Yang, Y., Huang, Z., Shen, H.T., and Hong, R. (December, January 28). Multiple Feature Hashing for Real-Time Large Scale Near-Duplicate Video Retrieval. Proceedings of the 19th ACM International Conference on Multimedia (ACM MM), Scottsdale, AZ, USA."},{"key":"ref_27","doi-asserted-by":"crossref","first-page":"1","DOI":"10.1145\/3356316","article-title":"Video retrieval with similarity-preserving deep temporal hashing","volume":"15","author":"Shen","year":"2019","journal-title":"ACM Trans. Multimedia Comput. Commun. Appl."},{"key":"ref_28","doi-asserted-by":"crossref","first-page":"102729","DOI":"10.1016\/j.dsp.2020.102729","article-title":"A supervised deep convolutional based bidirectional long short term memory video hashing for large scale video retrieval applications","volume":"102","author":"Anuranji","year":"2020","journal-title":"Digit. Signal Process."},{"key":"ref_29","doi-asserted-by":"crossref","first-page":"32","DOI":"10.1109\/TBDATA.2016.2530714","article-title":"Partial copy detection in videos: A benchmark and an evaluation of popular methods","volume":"2","author":"Jiang","year":"2016","journal-title":"IEEE Trans. Big Data"},{"key":"ref_30","doi-asserted-by":"crossref","unstructured":"Wang, L., Bao, Y., Li, H., Fan, X., and Luo, Z. (2017, January 4\u20136). Compact CNN based Video Representation for Efficient Video Copy Detection. Proceedings of the International Conference on Multimedia Modeling (MMM), Reykjavik, Iceland.","DOI":"10.1007\/978-3-319-51811-4_47"},{"key":"ref_31","doi-asserted-by":"crossref","unstructured":"Kordopatis-Zilos, G., Papadopoulos, S., Patras, I., and Kompatsiaris, Y. (2017, January 22\u201329). Near-Duplicate Video Retrieval with Deep Metric Learning. Proceedings of the 2017 IEEE International Conference on Computer Vision Workshops (ICCVW), Venice, Italy.","DOI":"10.1109\/ICCVW.2017.49"},{"key":"ref_32","doi-asserted-by":"crossref","first-page":"221","DOI":"10.1109\/TPAMI.2012.59","article-title":"3D convolutional neural networks for human action recognition","volume":"35","author":"Ji","year":"2012","journal-title":"IEEE Trans. Pattern Anal. Mach. Intell."},{"key":"ref_33","doi-asserted-by":"crossref","unstructured":"Huang, X., Shan, J., and Vaidya, V. (2017, January 18\u201321). Lung Nodule Detection in CT Using 3D Convolutional Neural Networks. Proceedings of the 2017 IEEE 14th International Symposium on Biomedical Imaging (ISBI), Melbourne, VIC, Australia.","DOI":"10.1109\/ISBI.2017.7950542"},{"key":"ref_34","doi-asserted-by":"crossref","unstructured":"Molchanov, P., Gupta, S., Kim, K., and Kautz, J. (2015, January 7\u201312). Hand Gesture Recognition with 3D Convolutional Neural Networks. Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition Workshops (CVPRW), Boston, MA, USA.","DOI":"10.1109\/CVPRW.2015.7301342"},{"key":"ref_35","doi-asserted-by":"crossref","unstructured":"Tran, D., Bourdev, L., Fergus, R., Torresani, L., and Paluri, M. (2015, January 13\u201316). Learning Spatiotemporal Features with 3d Convolutional Networks. Proceedings of the IEEE International Conference on Computer Vision (ICCV), Santiago, Chile.","DOI":"10.1109\/ICCV.2015.510"},{"key":"ref_36","doi-asserted-by":"crossref","unstructured":"Qiu, Z., Yao, T., and Mei, T. (2017, January 22\u201329). Learning Spatio-Temporal Representation with Pseudo-3d Residual Networks. Proceedings of the IEEE International Conference on Computer Vision (ICCV), Venice, Italy.","DOI":"10.1109\/ICCV.2017.590"},{"key":"ref_37","doi-asserted-by":"crossref","unstructured":"Carreira, J., and Zisserman, A. (2017, January 21\u201326). Quo Vadis, Action Recognition?. A New Model and the Kinetics Dataset. In proceedings of the IEEE Conference on Computer Vision and Pattern Recognition (CVPR), Honolulu, HI, USA.","DOI":"10.1109\/CVPR.2017.502"},{"key":"ref_38","doi-asserted-by":"crossref","unstructured":"Tran, D., Wang, H., Torresani, L., Ray, J., LeCun, Y., and Paluri, M. (2018, January 18\u201323). A Closer Look at Spatiotemporal Convolutions for Action Recognition. Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition (CVPR), Salt Lake City, UT, USA.","DOI":"10.1109\/CVPR.2018.00675"},{"key":"ref_39","unstructured":"Liu, W., Wang, J., Ji, R., Jiang, Y.G., and Chang, S.F. (2012, January 16\u201321). Supervised hashing with kernels. Proceedings of the 2012 IEEE Conference on Computer Vision and Pattern Recognition (CVPR), Providence, RI, USA."},{"key":"ref_40","unstructured":"Norouzi, M., and Fleet, D.J. (July, January 28). Minimal Loss Hashing for Compact Binary Codes. Proceedings of the 28th International Conference on Machine Learning (ICML), Bellevue, WA, USA."},{"key":"ref_41","unstructured":"Gionis, A., Indyk, P., and Motwani, R. (1999, January 7\u201310). Similarity Search in High Dimensions via Hashing. Proceedings of the 25th International Conference on Very Large Data Bases (VLDB), Edinburgh, Scotland, UK."},{"key":"ref_42","doi-asserted-by":"crossref","first-page":"2916","DOI":"10.1109\/TPAMI.2012.193","article-title":"Iterative quantization: A procrustean approach to learning binary codes for large-scale image retrieval","volume":"35","author":"Gong","year":"2013","journal-title":"IEEE Trans. Pattern Anal. Mach. Intell."},{"key":"ref_43","doi-asserted-by":"crossref","first-page":"3941","DOI":"10.1109\/TCYB.2016.2591068","article-title":"Unsupervised topic hypergraph hashing for efficient mobile image retrieval","volume":"47","author":"Zhu","year":"2016","journal-title":"IEEE Trans Cybern."},{"key":"ref_44","doi-asserted-by":"crossref","first-page":"472","DOI":"10.1109\/TKDE.2016.2562624","article-title":"Unsupervised visual hashing with semantic assistant for content-based image retrieval","volume":"29","author":"Zhu","year":"2016","journal-title":"IEEE Trans Knowl Data Eng."},{"key":"ref_45","doi-asserted-by":"crossref","first-page":"3210","DOI":"10.1109\/TIP.2018.2814344","article-title":"Self-supervised video hashing with hierarchical binary auto-encoder","volume":"27","author":"Song","year":"2018","journal-title":"IEEE Trans. Image Process."},{"key":"ref_46","doi-asserted-by":"crossref","first-page":"1993","DOI":"10.1109\/TIP.2018.2882155","article-title":"Unsupervised deep video hashing via balanced code for large-scale video retrieval","volume":"28","author":"Wu","year":"2018","journal-title":"IEEE Trans. Image Process."},{"key":"ref_47","doi-asserted-by":"crossref","unstructured":"Galanopoulos, D., and Mezaris, V. (2020, January 8\u201311). Attention Mechanisms, Signal Encodings and Fusion Strategies for Improved Ad-Hoc Video Search with Dual Encoding Networks. Proceedings of the 2020 International Conference on Multimedia Retrieval (ICMR), Dublin, Ireland.","DOI":"10.1145\/3372278.3390737"},{"key":"ref_48","doi-asserted-by":"crossref","unstructured":"Prathiba, T., and Kumari, R.S.S. (2020). Content based video retrieval system based on multimodal feature grouping by KFCM clustering algorithm to promote human\u2013computer interaction. J. Ambient Intell. Humaniz Comput., 1\u201315.","DOI":"10.1007\/s12652-020-02190-w"},{"key":"ref_49","unstructured":"Li, S., Chen, Z., Lu, J., Li, X., and Zhou, J. (November, January 27). Neighborhood Preserving Hashing for Scalable Video Retrieval. Proceedings of the IEEE International Conference on Computer Vision (ICCV), Seoul, Korea."},{"key":"ref_50","first-page":"2402","article-title":"Learning compact spatio-temporal features for fast content based video retrieval","volume":"9","author":"Kumar","year":"2019","journal-title":"Int. J. Innov. Tech. Explor. Eng."},{"key":"ref_51","doi-asserted-by":"crossref","unstructured":"Yuan, L., Wang, T., Zhang, X., Tay, F.E., Jie, Z., Liu, W., and Feng, J. (2020, January 13\u201319). Central Similarity Quantization for Efficient Image and Video Retrieval. Proceedings of the IEEE\/CVF Conference on Computer Vision and Pattern Recognition (CVPR), Seattle, WA, USA.","DOI":"10.1109\/CVPR42600.2020.00315"},{"key":"ref_52","doi-asserted-by":"crossref","unstructured":"Hara, K., Kataoka, H., and Satoh, Y. (2017, January 22\u201329). Learning Spatio-Temporal Features with 3D Residual Networks for Action Recognition. Proceedings of the IEEE International Conference on Computer Vision Workshops (ICCVW), Venice, Italy.","DOI":"10.1109\/ICCVW.2017.373"},{"key":"ref_53","doi-asserted-by":"crossref","unstructured":"He, K., Zhang, X., Ren, S., and Sun, J. (2016, January 27\u201330). Deep Residual Learning for Image Recognition. Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition (CVPR), Las Vegas, NV, USA.","DOI":"10.1109\/CVPR.2016.90"},{"key":"ref_54","unstructured":"Soomro, K., Zamir, A.R., and Shah, M. (2012). A dataset of 101 human action classes from videos in the wild. arXiv."},{"key":"ref_55","doi-asserted-by":"crossref","unstructured":"Jhuang, H., Gall, J., Zuffi, S., Schmid, C., and Black, M.J. (2013, January 1\u20138). Towards Understanding Action Recognition. Proceedings of the IEEE International Conference on Computer Vision (ICCV), Sydney, NSW, Australia.","DOI":"10.1109\/ICCV.2013.396"},{"key":"ref_56","doi-asserted-by":"crossref","unstructured":"Kuehne, H., Jhuang, H., Garrote, E., Poggio, T., and Serre, T. (2011, January 6\u201313). HMDB: A Large Video Database for Human Motion Recognition. Proceedings of the 2011 International Conference on Computer Vision (ICCV), Barcelona, Spain.","DOI":"10.1109\/ICCV.2011.6126543"},{"key":"ref_57","doi-asserted-by":"crossref","first-page":"1209","DOI":"10.1109\/TMM.2016.2645404","article-title":"Deep video hashing","volume":"19","author":"Liong","year":"2017","journal-title":"IEEE Trans. Multimed."}],"container-title":["Sensors"],"original-title":[],"language":"en","link":[{"URL":"https:\/\/www.mdpi.com\/1424-8220\/21\/9\/3094\/pdf","content-type":"unspecified","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2025,10,11]],"date-time":"2025-10-11T05:55:07Z","timestamp":1760162107000},"score":1,"resource":{"primary":{"URL":"https:\/\/www.mdpi.com\/1424-8220\/21\/9\/3094"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2021,4,29]]},"references-count":57,"journal-issue":{"issue":"9","published-online":{"date-parts":[[2021,5]]}},"alternative-id":["s21093094"],"URL":"https:\/\/doi.org\/10.3390\/s21093094","relation":{},"ISSN":["1424-8220"],"issn-type":[{"value":"1424-8220","type":"electronic"}],"subject":[],"published":{"date-parts":[[2021,4,29]]}}}