{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2026,6,13]],"date-time":"2026-06-13T06:29:30Z","timestamp":1781332170777,"version":"3.54.1"},"reference-count":31,"publisher":"Oxford University Press (OUP)","issue":"8","license":[{"start":{"date-parts":[[2024,3,23]],"date-time":"2024-03-23T00:00:00Z","timestamp":1711152000000},"content-version":"vor","delay-in-days":0,"URL":"https:\/\/academic.oup.com\/pages\/standard-publication-reuse-rights"}],"funder":[{"DOI":"10.13039\/501100001843","name":"Science and Engineering Research Board","doi-asserted-by":"publisher","id":[{"id":"10.13039\/501100001843","id-type":"DOI","asserted-by":"publisher"}]},{"DOI":"10.13039\/100021014","name":"Department of Science and Technology","doi-asserted-by":"publisher","award":["CRG\/2020\/001982"],"award-info":[{"award-number":["CRG\/2020\/001982"]}],"id":[{"id":"10.13039\/100021014","id-type":"DOI","asserted-by":"publisher"}]}],"content-domain":{"domain":[],"crossmark-restriction":false},"short-container-title":[],"published-print":{"date-parts":[[2024,8,11]]},"abstract":"<jats:title>Abstract<\/jats:title>\n               <jats:p>In this technological era, human activity recognition (HAR) plays a significant role in several applications like surveillance, health services, Internet of Things, etc. Recent advancements in deep learning and video summarization have motivated us to integrate these techniques for HAR. This paper introduces a computationally efficient HAR technique based on a deep learning framework, which works well in realistic and multi-view environments. Deep convolutional neural networks (DCNNs) normally suffer from different constraints, including data size dependencies, computational complexity, overfitting, training challenges and vanishing gradients. Additionally, with the use of advanced mobile vision devices, the demand for computationally efficient HAR algorithms with the requirement of limited computational resources is high. To address these issues, we used integration of DCNN with video summarization using keyframes. The proposed technique offers a solution that enhances performance with efficient resource utilization. For this, first, we designed a lightweight and computationally efficient deep learning architecture based on the concept of identity skip connections (features reusability), which preserves the gradient loss attenuation and can handle the enormous complexity of activity classes. Subsequently, we employed an efficient keyframe extraction technique to minimize redundancy and succinctly encapsulate the entire video content in a lesser number of frames. To evaluate the efficacy of the proposed method, we performed the experimentation on several publicly available datasets. The performance of the proposed method is measured in terms of evaluation parameters Precision, Recall, F-Measure and Classification Accuracy. The experimental results demonstrated the superiority of the presented algorithm over other existing state-of-the-art methods.<\/jats:p>","DOI":"10.1093\/comjnl\/bxae028","type":"journal-article","created":{"date-parts":[[2024,3,24]],"date-time":"2024-03-24T11:26:34Z","timestamp":1711279594000},"page":"2601-2609","source":"Crossref","is-referenced-by-count":6,"title":["Human Activity Recognition Based On Video Summarization And Deep Convolutional Neural Network"],"prefix":"10.1093","volume":"67","author":[{"given":"Arati","family":"Kushwaha","sequence":"first","affiliation":[{"name":"Department of Computer Engineering & Applications, GLA University , Mathura , India"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Manish","family":"Khare","sequence":"additional","affiliation":[{"name":"Dhirubhai Ambani Institute of Information and Communication Technology , Gandhinagar , India"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Reddy Mounika","family":"Bommisetty","sequence":"additional","affiliation":[{"name":"Department of Electronics & Communication, University of Allahabad , Prayagraj, Uttar Pradesh , India"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Ashish","family":"Khare","sequence":"additional","affiliation":[{"name":"Department of Computer Engineering & Applications, GLA University , Mathura , India"},{"name":"Department of Electronics & Communication, University of Allahabad , Prayagraj, Uttar Pradesh , India"}],"role":[{"vocabulary":"crossref","role":"author"}]}],"member":"286","published-online":{"date-parts":[[2024,3,23]]},"reference":[{"key":"2024081612055901300_ref1","doi-asserted-by":"crossref","first-page":"1366","DOI":"10.1007\/s11263-022-01594-9","article-title":"Human action recognition and prediction: a survey","volume":"130","author":"Kong","year":"2022","journal-title":"Int. J. Comput. Vision"},{"key":"2024081612055901300_ref2","first-page":"3200","article-title":"Human action recognition from various data modalities: a review","volume":"45","author":"Sun","year":"2022","journal-title":"IEEE Trans. Pattern Anal. Mach. Intell."},{"key":"2024081612055901300_ref3","doi-asserted-by":"crossref","first-page":"2259","DOI":"10.1007\/s10462-020-09904-8","article-title":"A survey on video-based human action recognition: recent updates, datasets, challenges, and applications","volume":"54","author":"Pareek","year":"2021","journal-title":"Artif. Intell. Rev."},{"key":"2024081612055901300_ref4","doi-asserted-by":"crossref","first-page":"1","DOI":"10.1016\/j.patcog.2018.07.028","article-title":"Asymmetric 3d convolutional neural networks for action recognition","volume":"85","author":"Yang","year":"2019","journal-title":"Pattern Recognit."},{"key":"2024081612055901300_ref5","doi-asserted-by":"crossref","first-page":"13321","DOI":"10.1007\/s00521-023-08440-0","article-title":"Micro-network-based deep convolutional neural network for human activity recognition from realistic and multi-view visual data","volume":"35","author":"Kushwaha","year":"2023","journal-title":"Neural Computi. Appl."},{"key":"2024081612055901300_ref6","doi-asserted-by":"crossref","first-page":"2250045","DOI":"10.1142\/S0218213022500452","article-title":"Micro-network based convolutional neural network with integration of multilayer feature fusion strategy for human activity recognition","volume":"31","author":"Kushwaha","year":"2022","journal-title":"Int. J. Artif. Intell. Tools"},{"key":"2024081612055901300_ref7","doi-asserted-by":"crossref","first-page":"267","DOI":"10.1007\/s00530-019-00642-8","article-title":"Keyframe extraction using Pearson correlation coefficient and color moments","volume":"26","author":"Bommisetty","year":"2020","journal-title":"Multimedia Syst."},{"key":"2024081612055901300_ref8","doi-asserted-by":"crossref","first-page":"84","DOI":"10.1145\/3065386","article-title":"Imagenet classification with deep convolutional neural networks","volume":"60","author":"Krizhevsky","year":"2017","journal-title":"Commun. ACM"},{"key":"2024081612055901300_ref9","first-page":"1","article-title":"Very deep convolutional networks for large-scale image recognition","volume":"6","author":"Simonyan","year":"2014","journal-title":"Computer Vision and Pattern Recognition"},{"key":"2024081612055901300_ref10","first-page":"1","article-title":"Going deeper with convolutions","volume-title":"Proc. of the IEEE Conf. on Computer Vision and Pattern Recognition (CVPR)","author":"Szegedy","year":"2015"},{"key":"2024081612055901300_ref11","doi-asserted-by":"crossref","first-page":"249","DOI":"10.1016\/j.cviu.2006.07.013","article-title":"Free viewpoint action recognition using motion history volumes","volume":"104","author":"Weinland","year":"2006","journal-title":"Comput. Vision Image Understanding"},{"key":"2024081612055901300_ref12","first-page":"2556","article-title":"Hmdb: a large video database for human motion recognition","volume-title":"Proc. of the Int. Conference on Computer Vision (ICCV)","author":"Kuehne","year":"2011"},{"key":"2024081612055901300_ref13","first-page":"1","article-title":"An end-to-end generative framework for video segmentation and recognition","volume-title":"Proc. of the IEEE Winter Conference on Applications of Computer Vision (WACV)","author":"Kuehne","year":"2016"},{"key":"2024081612055901300_ref14","first-page":"1","article-title":"Youtube-8m: a large-scale video classification benchmark","volume":"1","author":"Abu-El-Haija","year":"2016","journal-title":"Computer Vision and Pattern Recognition"},{"key":"2024081612055901300_ref15","first-page":"1","article-title":"The kinetics human action video dataset","volume":"1","author":"Kay","year":"2017","journal-title":"Computer Vision and Pattern Recognition"},{"key":"2024081612055901300_ref16","doi-asserted-by":"crossref","first-page":"32511","DOI":"10.1007\/s11042-021-11207-1","article-title":"On integration of multiple features for human activity recognition in video sequences","volume":"80","author":"Kushwaha","year":"2021","journal-title":"Multimed. Tools Appl."},{"key":"2024081612055901300_ref17","doi-asserted-by":"crossref","first-page":"281","DOI":"10.1007\/s10044-019-00789-0","article-title":"Human action recognition: a framework of statistical weighted segmentation and rank correlation-based selection","volume":"23","author":"Sharif","year":"2020","journal-title":"Pattern Anal. Appl."},{"key":"2024081612055901300_ref18","doi-asserted-by":"crossref","first-page":"2250009","DOI":"10.1142\/S0219467822500097","article-title":"Human activity recognition algorithm in video sequences based on integration of magnitude and orientation information of optical flow","volume":"22","author":"Kushwaha","year":"2022","journal-title":"Int. J. Image Graphics"},{"key":"2024081612055901300_ref19","doi-asserted-by":"crossref","first-page":"e7571","DOI":"10.1002\/cpe.7571","article-title":"Human activity recognition based on integration of multilayer information of convolutional neural network architecture","volume":"35","author":"Kushwaha","year":"2023","journal-title":"Concurrency Comput. Pract. Exper."},{"key":"2024081612055901300_ref20","doi-asserted-by":"crossref","first-page":"690","DOI":"10.1007\/s10489-020-01823-z","article-title":"A combined multiple action recognition and summarization for surveillance video sequences","volume":"51","author":"Elharrouss","year":"2021","journal-title":"Appl. Intell."},{"key":"2024081612055901300_ref21","doi-asserted-by":"crossref","first-page":"2589","DOI":"10.1007\/s10489-020-01905-y","article-title":"Video sketch: a middle-level representation for action recognition","volume":"51","author":"Zhang","year":"2021","journal-title":"Appl. Intell."},{"key":"2024081612055901300_ref22","first-page":"3333","article-title":"Multiview transformers for video recognition","volume-title":"Proc. of the IEEE\/CVF Conference on Computer Vision and Pattern Recognition (CVPR)","author":"Yan","year":"2022"},{"key":"2024081612055901300_ref23","first-page":"770","article-title":"Deep residual learning for image recognition","volume-title":"Proc. of the IEEE Conf. on Computer Vision and Pattern Recognition (CVPR)","author":"He","year":"2016"},{"key":"2024081612055901300_ref24","doi-asserted-by":"crossref","first-page":"130","DOI":"10.3390\/jimaging9070130","article-title":"Human activity recognition using cascaded dual attention cnn and bi-directional gru framework","volume":"9","author":"Ullah","year":"2023","journal-title":"J. Imaging"},{"key":"2024081612055901300_ref25","doi-asserted-by":"crossref","first-page":"191997","DOI":"10.1109\/ACCESS.2020.3033190","article-title":"Video activity recognition with varying rhythms","volume":"8","author":"Ayhan","year":"2020","journal-title":"IEEE Access"},{"key":"2024081612055901300_ref26","first-page":"19880","article-title":"Bridge-prompt: Towards ordinal action understanding in instructional videos","volume-title":"Proc. of the IEEE\/CVF Conf. on Computer Vision and Pattern Recognition (CVPR)","author":"Li","year":"2022"},{"key":"2024081612055901300_ref27","first-page":"279","article-title":"A generalized and robust framework for timestamp supervision in temporal action segmentation","volume-title":"Proc. of the European Conference on Computer Vision","author":"Rahaman","year":"2022"},{"key":"2024081612055901300_ref28","first-page":"8365","article-title":"Temporal action segmentation from timestamp supervision","volume-title":"Proceedings of the IEEE\/CVF conference on computer vision and pattern recognition (CVPR)","author":"Li","year":"2021"},{"key":"2024081612055901300_ref29","first-page":"6807","article-title":"Large scale video representation learning via relational graph clustering","volume-title":"Proc. of the IEEE\/CVF Conf. on Computer Vision and Pattern Recognition (CVPR)","author":"Lee","year":"2020"},{"key":"2024081612055901300_ref30","doi-asserted-by":"crossref","first-page":"948","DOI":"10.3390\/app14020948","article-title":"Sports video classification method based on improved deep learning","volume":"14","author":"Gao","year":"2024","journal-title":"Appl. Sci."},{"key":"2024081612055901300_ref31","first-page":"1","article-title":"Actionhub: a large-scale action video description dataset for zero-shot action recognition","author":"Zhou","year":"2024","journal-title":"arXiv preprint arXiv:2401.11654, NA"}],"container-title":["The Computer Journal"],"original-title":[],"language":"en","link":[{"URL":"https:\/\/academic.oup.com\/comjnl\/article-pdf\/67\/8\/2601\/58796446\/bxae028.pdf","content-type":"application\/pdf","content-version":"vor","intended-application":"syndication"},{"URL":"https:\/\/academic.oup.com\/comjnl\/article-pdf\/67\/8\/2601\/58796446\/bxae028.pdf","content-type":"unspecified","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2024,8,16]],"date-time":"2024-08-16T12:07:24Z","timestamp":1723810044000},"score":1,"resource":{"primary":{"URL":"https:\/\/academic.oup.com\/comjnl\/article\/67\/8\/2601\/7634135"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2024,3,23]]},"references-count":31,"journal-issue":{"issue":"8","published-online":{"date-parts":[[2024,3,23]]},"published-print":{"date-parts":[[2024,8,11]]}},"URL":"https:\/\/doi.org\/10.1093\/comjnl\/bxae028","relation":{},"ISSN":["0010-4620","1460-2067"],"issn-type":[{"value":"0010-4620","type":"print"},{"value":"1460-2067","type":"electronic"}],"subject":[],"published-other":{"date-parts":[[2024,8]]},"published":{"date-parts":[[2024,3,23]]}}}