{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2026,4,22]],"date-time":"2026-04-22T05:35:42Z","timestamp":1776836142491,"version":"3.51.2"},"reference-count":49,"publisher":"Wiley","license":[{"start":{"date-parts":[[2020,12,16]],"date-time":"2020-12-16T00:00:00Z","timestamp":1608076800000},"content-version":"unspecified","delay-in-days":0,"URL":"https:\/\/creativecommons.org\/licenses\/by\/4.0\/"}],"funder":[{"name":"NVIDIA Corporation"}],"content-domain":{"domain":[],"crossmark-restriction":false},"short-container-title":["Journal of Sensors"],"published-print":{"date-parts":[[2020,12,16]]},"abstract":"<jats:p>Head detection in real-world videos is a classical research problem in computer vision. Head detection in videos is challenging than in a single image due to many nuisances that are commonly observed in natural videos, including arbitrary poses, appearances, and scales. Generally, head detection is treated as a particular case of object detection in a single image. However, the performance of object detectors deteriorates in unconstrained videos. In this paper, we propose a temporal consistency model (TCM) to enhance the performance of a generic object detector by integrating spatial-temporal information that exists among subsequent frames of a particular video. Generally, our model takes detection from a generic detector as input and improves mean average precision (mAP) by recovering missed detection and suppressing false positives. We compare and evaluate the proposed framework on four challenging datasets, i.e., HollywoodHeads, Casablanca, BOSS, and PAMELA. Experimental evaluation shows that the performance is improved by employing the proposed TCM model. We demonstrate both qualitatively and quantitatively that our proposed framework obtains significant improvements over other methods.<\/jats:p>","DOI":"10.1155\/2020\/8861296","type":"journal-article","created":{"date-parts":[[2020,12,17]],"date-time":"2020-12-17T02:19:37Z","timestamp":1608171577000},"page":"1-13","source":"Crossref","is-referenced-by-count":13,"title":["TCM: Temporal Consistency Model for Head Detection in Complex Videos"],"prefix":"10.1155","volume":"2020","author":[{"ORCID":"https:\/\/orcid.org\/0000-0002-7406-8441","authenticated-orcid":true,"given":"Sultan Daud","family":"Khan","sequence":"first","affiliation":[{"name":"Department of Computer Science, National University of Technology, Pakistan"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Ahmed B.","family":"Altamimi","sequence":"additional","affiliation":[{"name":"Department of Computer Science and Software Engineering, University of Ha\u2019il, Saudi Arabia"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"ORCID":"https:\/\/orcid.org\/0000-0002-0222-6340","authenticated-orcid":true,"given":"Mohib","family":"Ullah","sequence":"additional","affiliation":[{"name":"Department of Computer Science, Norwegian University of Science and Technology, Norway"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"ORCID":"https:\/\/orcid.org\/0000-0002-2434-0849","authenticated-orcid":true,"given":"Habib","family":"Ullah","sequence":"additional","affiliation":[{"name":"Department of Computer Science and Software Engineering, University of Ha\u2019il, Saudi Arabia"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"ORCID":"https:\/\/orcid.org\/0000-0002-4823-5250","authenticated-orcid":true,"given":"Faouzi Alaya","family":"Cheikh","sequence":"additional","affiliation":[{"name":"Department of Computer Science, Norwegian University of Science and Technology, Norway"}],"role":[{"role":"author","vocabulary":"crossref"}]}],"member":"311","reference":[{"issue":"9","key":"1","doi-asserted-by":"crossref","first-page":"2146","DOI":"10.1109\/TPAMI.2018.2849374","article-title":"On detection, data association and segmentation for multi-target tracking","volume":"41","author":"Y. Tian","year":"2018","journal-title":"IEEE Transactions on Pattern Analysis and Machine Intelligence"},{"key":"2","first-page":"1816","article-title":"A directed sparse graphical model for multi-target tracking","author":"M. Ullah"},{"key":"3","first-page":"6479","article-title":"Real-world anomaly detection in surveillance videos","author":"W. Sultani"},{"key":"4","doi-asserted-by":"publisher","DOI":"10.1117\/12.2040521"},{"key":"5","doi-asserted-by":"publisher","DOI":"10.1016\/j.engappai.2019.07.009"},{"key":"6","doi-asserted-by":"publisher","DOI":"10.1080\/15472450.2020.1746909"},{"key":"7","doi-asserted-by":"publisher","DOI":"10.1016\/j.neucom.2015.11.049"},{"key":"8","doi-asserted-by":"publisher","DOI":"10.1109\/ICIP.2016.7532547"},{"key":"9","doi-asserted-by":"publisher","DOI":"10.1109\/TCSVT.2019.2890840"},{"key":"10","first-page":"3431","article-title":"Fully convolutional networks for semantic segmentation","author":"J. Long"},{"key":"11","doi-asserted-by":"publisher","DOI":"10.1109\/ICIP.2018.8451653"},{"key":"12","doi-asserted-by":"publisher","DOI":"10.1109\/cvpr.2014.81"},{"key":"13","first-page":"21","article-title":"Ssd: single shot multibox detector","author":"W. Liu"},{"key":"14","doi-asserted-by":"publisher","DOI":"10.1109\/cvpr.2017.690"},{"key":"15","first-page":"91","article-title":"Faster r-cnn: towards real-time object detection with region proposal networks","author":"S. Ren"},{"key":"16","first-page":"2893","article-title":"Context-aware cnns for person head detection","author":"T.-H. Vu"},{"key":"17","doi-asserted-by":"publisher","DOI":"10.1007\/s11263-013-0620-5"},{"key":"18","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR.2008.4587533"},{"key":"19"},{"key":"20","article-title":"UCL Transport Institute"},{"key":"21","doi-asserted-by":"publisher","DOI":"10.1109\/tcsvt.2017.2711015"},{"issue":"8","key":"22","doi-asserted-by":"crossref","first-page":"1875","DOI":"10.1109\/TCSVT.2017.2691801","article-title":"Local large-margin multimetric learning for face and kinship verification","volume":"28","author":"J. Hu","year":"2017","journal-title":"IEEE Transactions on Circuits and Systems for Video Technology"},{"issue":"3","key":"23","doi-asserted-by":"crossref","first-page":"529","DOI":"10.1109\/TCSVT.2015.2412831","article-title":"Localized multi-feature metric learning for image-set-based face recognition","volume":"26","author":"J. Lu","year":"2015","journal-title":"IEEE Transactions on Circuits and Systems for Video Technology"},{"issue":"8","key":"24","first-page":"3636","article-title":"Learning rotation-invariant local binary descriptor","volume":"26","author":"Y. Duan","year":"2017","journal-title":"IEEE Transactions on Image Processing"},{"key":"25","doi-asserted-by":"publisher","DOI":"10.1109\/TPAMI.2015.2408359"},{"key":"26","doi-asserted-by":"publisher","DOI":"10.1109\/TIP.2017.2713940"},{"key":"27","first-page":"1891","article-title":"Deep learning face representation from predicting 10,000 classes","author":"Y. Sun"},{"key":"28","first-page":"5325","article-title":"A convolutional neural network cascade for face detection","author":"H. Li"},{"key":"29","first-page":"3676","article-title":"From facial parts responses to face detection: a deep learning approach","author":"S. Yang"},{"key":"30","first-page":"3171","article-title":"Detecting faces using inside cascaded contextual cnn","author":"K. Zhang"},{"key":"31","doi-asserted-by":"crossref","first-page":"57","DOI":"10.1007\/978-3-319-61657-5_3","article-title":"Cms-rcnn: contextual multi-scale region-based cnn for unconstrained face detection","volume-title":"Deep Learning for Biometrics","author":"C. Zhu","year":"2017"},{"key":"32","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR.2017.166"},{"key":"33","first-page":"6186","article-title":"Scale-aware face detection","author":"Z. Hao"},{"key":"34","doi-asserted-by":"publisher","DOI":"10.1109\/TPAMI.2017.2710183"},{"key":"35","doi-asserted-by":"publisher","DOI":"10.1023\/B:VISI.0000013087.49260.fb"},{"key":"36","first-page":"2497","article-title":"The fastest deformable part model for object detection","author":"J. Yan"},{"key":"37","doi-asserted-by":"publisher","DOI":"10.1109\/ICIP.2016.7532426"},{"key":"38","doi-asserted-by":"publisher","DOI":"10.1109\/ROBIO.2017.8324433"},{"key":"39","doi-asserted-by":"publisher","DOI":"10.1109\/FG.2018.00089"},{"key":"40","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR.2009.5206848"},{"key":"41","article-title":"Very deep convolutional networks for large-scale image recognition","author":"K. Simonyan","year":"2014"},{"key":"42","first-page":"818","article-title":"Visualizing and understanding convolutional networks","author":"M. D. Zeiler"},{"key":"43","doi-asserted-by":"publisher","DOI":"10.1016\/0004-3702(81)90024-2"},{"key":"44","doi-asserted-by":"publisher","DOI":"10.1007\/BF00202895"},{"key":"45","doi-asserted-by":"publisher","DOI":"10.1109\/ICMA.2007.4303646"},{"key":"46","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR.2005.177"},{"key":"47","doi-asserted-by":"publisher","DOI":"10.1109\/ICPR.2000.902888"},{"key":"48","article-title":"Fchd: fast and accurate head detection in crowded scenes","author":"A. Vora","year":"2018"},{"key":"49","first-page":"2325","article-title":"End-to-end people detection in crowded scenes","author":"R. Stewart"}],"container-title":["Journal of Sensors"],"original-title":[],"language":"en","link":[{"URL":"http:\/\/downloads.hindawi.com\/journals\/js\/2020\/8861296.pdf","content-type":"application\/pdf","content-version":"vor","intended-application":"text-mining"},{"URL":"http:\/\/downloads.hindawi.com\/journals\/js\/2020\/8861296.xml","content-type":"application\/xml","content-version":"vor","intended-application":"text-mining"},{"URL":"http:\/\/downloads.hindawi.com\/journals\/js\/2020\/8861296.pdf","content-type":"unspecified","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2020,12,17]],"date-time":"2020-12-17T02:19:44Z","timestamp":1608171584000},"score":1,"resource":{"primary":{"URL":"https:\/\/www.hindawi.com\/journals\/js\/2020\/8861296\/"}},"subtitle":[],"editor":[{"given":"Abdellah","family":"Touhafi","sequence":"additional","affiliation":[],"role":[{"role":"editor","vocabulary":"crossref"}]}],"short-title":[],"issued":{"date-parts":[[2020,12,16]]},"references-count":49,"alternative-id":["8861296","8861296"],"URL":"https:\/\/doi.org\/10.1155\/2020\/8861296","relation":{},"ISSN":["1687-7268","1687-725X"],"issn-type":[{"value":"1687-7268","type":"electronic"},{"value":"1687-725X","type":"print"}],"subject":[],"published":{"date-parts":[[2020,12,16]]}}}