{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2026,4,16]],"date-time":"2026-04-16T15:13:01Z","timestamp":1776352381911,"version":"3.51.2"},"reference-count":44,"publisher":"MDPI AG","issue":"3","license":[{"start":{"date-parts":[[2021,3,13]],"date-time":"2021-03-13T00:00:00Z","timestamp":1615593600000},"content-version":"vor","delay-in-days":0,"URL":"https:\/\/creativecommons.org\/licenses\/by\/4.0\/"}],"content-domain":{"domain":[],"crossmark-restriction":false},"short-container-title":["J. Imaging"],"abstract":"<jats:p>Methods for No-Reference Video Quality Assessment (NR-VQA) of consumer-produced video content are largely investigated due to the spread of databases containing videos affected by natural distortions. In this work, we design an effective and efficient method for NR-VQA. The proposed method exploits a novel sampling module capable of selecting a predetermined number of frames from the whole video sequence on which to base the quality assessment. It encodes both the quality attributes and semantic content of video frames using two lightweight Convolutional Neural Networks (CNNs). Then, it estimates the quality score of the entire video using a Support Vector Regressor (SVR). We compare the proposed method against several relevant state-of-the-art methods using four benchmark databases containing user generated videos (CVD2014, KoNViD-1k, LIVE-Qualcomm, and LIVE-VQC). The results show that the proposed method at a substantially lower computational cost predicts subjective video quality in line with the state of the art methods on individual databases and generalizes better than existing methods in cross-database setup.<\/jats:p>","DOI":"10.3390\/jimaging7030055","type":"journal-article","created":{"date-parts":[[2021,3,14]],"date-time":"2021-03-14T22:13:10Z","timestamp":1615759990000},"page":"55","update-policy":"https:\/\/doi.org\/10.3390\/mdpi_crossmark_policy","source":"Crossref","is-referenced-by-count":13,"title":["An Efficient Method for No-Reference Video Quality Assessment"],"prefix":"10.3390","volume":"7","author":[{"ORCID":"https:\/\/orcid.org\/0000-0002-0009-8007","authenticated-orcid":false,"given":"Mirko","family":"Agarla","sequence":"first","affiliation":[{"name":"Department of Informatics, Systems and Communication, University of Milano-Bicocca, Viale Sarca, 336, 20126 Milano, Italy"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"ORCID":"https:\/\/orcid.org\/0000-0002-5925-2646","authenticated-orcid":false,"given":"Luigi","family":"Celona","sequence":"additional","affiliation":[{"name":"Department of Informatics, Systems and Communication, University of Milano-Bicocca, Viale Sarca, 336, 20126 Milano, Italy"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"ORCID":"https:\/\/orcid.org\/0000-0001-7461-1451","authenticated-orcid":false,"given":"Raimondo","family":"Schettini","sequence":"additional","affiliation":[{"name":"Department of Informatics, Systems and Communication, University of Milano-Bicocca, Viale Sarca, 336, 20126 Milano, Italy"}],"role":[{"role":"author","vocabulary":"crossref"}]}],"member":"1968","published-online":{"date-parts":[[2021,3,13]]},"reference":[{"key":"ref_1","doi-asserted-by":"crossref","unstructured":"Li, Y., Meng, S., Zhang, X., Wang, S., Wang, Y., and Ma, S. (2020, January 6\u20138). UGC-VIDEO: Perceptual quality assessment of user-generated videos. Proceedings of the IEEE Conference on Multimedia Information Processing and Retrieval (MIPR), Shenzhen, China.","DOI":"10.1109\/MIPR49039.2020.00015"},{"key":"ref_2","doi-asserted-by":"crossref","unstructured":"Yim, J.G., Wang, Y., Birkbeck, N., and Adsumilli, B. (2020, January 25\u201328). Subjective quality assessment for youtube ugc dataset. Proceedings of the IEEE International Conference on Image Processing (ICIP), Abu Dhabi, United Arab Emirates.","DOI":"10.1109\/ICIP40778.2020.9191194"},{"key":"ref_3","first-page":"113160P","article-title":"Towards a video quality assessment based framework for enhancement of laparoscopic videos","volume":"Volume 11316","author":"Khan","year":"2020","journal-title":"Medical Imaging 2020: Image Perception, Observer Performance, and Technology Assessment"},{"key":"ref_4","doi-asserted-by":"crossref","first-page":"5923","DOI":"10.1109\/TIP.2019.2923051","article-title":"Two-Level Approach for No-Reference Consumer Video Quality Assessment","volume":"28","author":"Korhonen","year":"2019","journal-title":"IEEE Trans. Image Process."},{"key":"ref_5","doi-asserted-by":"crossref","unstructured":"Agarla, M., Celona, L., and Schettini, R. (2020). No-Reference Quality Assessment of In-Capture Distorted Videos. MDPI J. Imaging, 6.","DOI":"10.3390\/jimaging6080074"},{"key":"ref_6","unstructured":"Hosu, V., Hahn, F., Jenadeleh, M., Lin, H., Men, H., Szir\u00e1nyi, T., Li, S., and Saupe, D. (June, January 31). The Konstanz natural video database (KoNViD-1k). Proceedings of the International Conference on Quality of Multimedia Experience (QoMEX), Erfurt, Germany."},{"key":"ref_7","doi-asserted-by":"crossref","first-page":"612","DOI":"10.1109\/TIP.2018.2869673","article-title":"Large-scale study of perceptual video quality","volume":"28","author":"Sinno","year":"2018","journal-title":"IEEE Trans. Image Process."},{"key":"ref_8","doi-asserted-by":"crossref","first-page":"2061","DOI":"10.1109\/TCSVT.2017.2707479","article-title":"In-capture mobile video distortions: A study of subjective behavior and objective algorithms","volume":"28","author":"Ghadiyaram","year":"2017","journal-title":"IEEE Trans. Circuits Syst. Video Technol."},{"key":"ref_9","doi-asserted-by":"crossref","unstructured":"Li, D., Jiang, T., and Jiang, M. (2019, January 21\u201325). Quality Assessment of In-the-Wild Videos. Proceedings of the ACM International Conference on Multimedia, Nice, France.","DOI":"10.1145\/3343031.3351028"},{"key":"ref_10","doi-asserted-by":"crossref","unstructured":"He, K., Zhang, X., Ren, S., and Sun, J. (2016, January 27\u201330). Deep residual learning for image recognition. Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition (CVPR), Las Vegas, NV, USA.","DOI":"10.1109\/CVPR.2016.90"},{"key":"ref_11","doi-asserted-by":"crossref","unstructured":"Sandler, M., Howard, A., Zhu, M., Zhmoginov, A., and Chen, L.C. (2018, January 18\u201323). Mobilenetv2: Inverted residuals and linear bottlenecks. Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition, Salt Lake City, UT, USA.","DOI":"10.1109\/CVPR.2018.00474"},{"key":"ref_12","doi-asserted-by":"crossref","first-page":"209","DOI":"10.1109\/LSP.2012.2227726","article-title":"Making a \u201cCompletely Blind\u201d Image Quality Analyzer","volume":"20","author":"Mittal","year":"2013","journal-title":"IEEE Signal Process. Lett."},{"key":"ref_13","doi-asserted-by":"crossref","first-page":"4695","DOI":"10.1109\/TIP.2012.2214050","article-title":"No-Reference Image Quality Assessment in the Spatial Domain","volume":"21","author":"Mittal","year":"2012","journal-title":"IEEE Trans. Image Process."},{"key":"ref_14","first-page":"32","article-title":"Perceptual Quality Prediction on Authentically Distorted Images Using a Bag of Features Approach","volume":"17","author":"Ghadiyaram","year":"2017","journal-title":"Assoc. Res. Vis. Ophthalmol. J. Vis."},{"key":"ref_15","doi-asserted-by":"crossref","first-page":"2957","DOI":"10.1109\/TIP.2017.2685941","article-title":"No-Reference Quality Assessment of Tone-Mapped HDR Pictures","volume":"26","author":"Kundu","year":"2017","journal-title":"IEEE Trans. Image Process."},{"key":"ref_16","doi-asserted-by":"crossref","unstructured":"Saad, M., and Bovik, A. (2012, January 4\u20137). Blind quality assessment of videos using a model of natural scene statistics and motion coherency. Proceedings of the Asilomar Conference on Signals, Systems and Computers (ASILOMAR), Pacific Grove, CA, USA.","DOI":"10.1109\/ACSSC.2012.6489018"},{"key":"ref_17","doi-asserted-by":"crossref","unstructured":"Xu, J., Ye, P., Liu, Y., and Doermann, D. (2014, January 27\u201330). No-reference video quality assessment via feature learning. Proceedings of the International Conference on Image Processing (ICIP), Paris, France.","DOI":"10.1109\/ICIP.2014.7025098"},{"key":"ref_18","doi-asserted-by":"crossref","first-page":"3073","DOI":"10.1109\/TIP.2016.2562513","article-title":"CVD2014\u2014A database for evaluating no-reference video quality assessment algorithms","volume":"25","author":"Nuutinen","year":"2016","journal-title":"IEEE Trans. Image Process."},{"key":"ref_19","doi-asserted-by":"crossref","first-page":"5612","DOI":"10.1109\/TIP.2020.2984879","article-title":"No-Reference Video Quality Assessment Using Natural Spatiotemporal Scene Statistics","volume":"29","author":"Dendi","year":"2020","journal-title":"IEEE Trans. Image Process."},{"key":"ref_20","doi-asserted-by":"crossref","unstructured":"Ebenezer, J.P., Shang, Z., Wu, Y., Wei, H., and Bovik, A.C. (2020, January 21\u201324). No-Reference Video Quality Assessment Using Space-Time Chips. Proceedings of the IEEE International Workshop on Multimedia Signal Processing (MMSP), Tampere, Finland.","DOI":"10.1109\/MMSP48831.2020.9287151"},{"key":"ref_21","doi-asserted-by":"crossref","unstructured":"Korhonen, J., Su, Y., and You, J. (2020, January 12\u201316). Blind Natural Video Quality Prediction via Statistical Temporal Features and Deep Spatial Features. Proceedings of the ACM 28th ACM International Conference on Multimedia, Seattle, WA, USA.","DOI":"10.1145\/3394171.3413845"},{"key":"ref_22","doi-asserted-by":"crossref","first-page":"1044","DOI":"10.1109\/TCSVT.2015.2430711","article-title":"No-Reference Video Quality Assessment With 3D Shearlet Transform and Convolutional Neural Networks","volume":"26","author":"Li","year":"2016","journal-title":"IEEE Trans. Circuits Syst. Video Technol."},{"key":"ref_23","doi-asserted-by":"crossref","unstructured":"Wang, C., Su, L., and Zhang, W. (2018, January 10\u201312). COME for No-Reference Video Quality Assessment. Proceedings of the IEEE Conference on Multimedia Information Processing and Retrieval (MIPR), Miami, FL, USA.","DOI":"10.1109\/MIPR.2018.00056"},{"key":"ref_24","doi-asserted-by":"crossref","first-page":"011006","DOI":"10.1117\/1.3267105","article-title":"Most apparent distortion: Full-reference image quality assessment and the role of strategy","volume":"19","author":"Larson","year":"2010","journal-title":"J. Electron. Imaging"},{"key":"ref_25","doi-asserted-by":"crossref","unstructured":"Deng, J., Dong, W., Socher, R., Li, L.J., Li, K., and Fei-Fei, L. (2009, January 20\u201325). Imagenet: A large-scale hierarchical image database. Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition (CVPR), Miami, FL, USA.","DOI":"10.1109\/CVPR.2009.5206848"},{"key":"ref_26","doi-asserted-by":"crossref","unstructured":"Cho, K., van Merri\u00ebnboer, B., Gulcehre, C., Bahdanau, D., Bougares, F., Schwenk, H., and Bengio, Y. (2014, January 25\u201329). Learning Phrase Representations using RNN Encoder\u2013Decoder for Statistical Machine Translation. Proceedings of the Conference on Empirical Methods in Natural Language Processing (EMNLP), Doha, Qatar.","DOI":"10.3115\/v1\/D14-1179"},{"key":"ref_27","doi-asserted-by":"crossref","unstructured":"Chen, P., Li, L., Ma, L., Wu, J., and Shi, G. (2020, January 26\u201329). RIRNet: Recurrent-In-Recurrent Network for Video Quality Assessment. Proceedings of the ACM International Conference on Multimedia, Dublin, Ireland.","DOI":"10.1145\/3394171.3413717"},{"key":"ref_28","doi-asserted-by":"crossref","unstructured":"Li, D., Jiang, T., and Jiang, M. (2021). Unified Quality Assessment of in-the-Wild Videos with Mixed Datasets Training. Int. J. Comput. Vis.","DOI":"10.1007\/s11263-020-01408-w"},{"key":"ref_29","doi-asserted-by":"crossref","unstructured":"Tang, Y., Jiang, S., Xu, S., Liu, T., and Li, C. (2019). Blind Image Quality Assessment Based on Multi-Window Method and HSV Color Space. Appl. Sci., 9.","DOI":"10.3390\/app9122499"},{"key":"ref_30","first-page":"105","article-title":"Quality Measurement of Blurred Images Using NMSE and SSIM Metrics in HSV and RGB Color Spaces","volume":"1","year":"2015","journal-title":"Phys. J."},{"key":"ref_31","doi-asserted-by":"crossref","first-page":"77440Z","DOI":"10.1117\/12.862431","article-title":"Image quality assessment and human visual system","volume":"Volume 7744","author":"Gao","year":"2010","journal-title":"Visual Communications and Image Processing 2010"},{"key":"ref_32","doi-asserted-by":"crossref","first-page":"355","DOI":"10.1007\/s11760-017-1166-8","article-title":"On the use of deep learning for blind image quality assessment","volume":"12","author":"Bianco","year":"2017","journal-title":"Signal Image Video Process."},{"key":"ref_33","doi-asserted-by":"crossref","first-page":"237","DOI":"10.1016\/j.image.2017.10.009","article-title":"Semantic-aware blind image quality assessment","volume":"60","author":"Siahaan","year":"2018","journal-title":"Signal Process. Image Commun."},{"key":"ref_34","doi-asserted-by":"crossref","first-page":"195","DOI":"10.1016\/j.jvcir.2018.11.038","article-title":"Blind image quality assessment with semantic information","volume":"58","author":"Ji","year":"2019","journal-title":"J. Vis. Commun. Image Represent."},{"key":"ref_35","doi-asserted-by":"crossref","first-page":"64270","DOI":"10.1109\/ACCESS.2018.2877890","article-title":"Benchmark Analysis of Representative Deep Neural Network Architectures","volume":"6","author":"Bianco","year":"2018","journal-title":"IEEE Access"},{"key":"ref_36","doi-asserted-by":"crossref","unstructured":"Virtanen, T., Nuutinen, M., Vaahteranoksa, M., Oittinen, P., and H\u00e4kkinen, J. (2014). CID2013: A Database for Evaluating No-Reference Image Quality Assessment Algorithms. IEEE Trans. Image Process., 24.","DOI":"10.1109\/TIP.2014.2378061"},{"key":"ref_37","doi-asserted-by":"crossref","unstructured":"Fang, Y., Zhu, H., Zeng, Y., Ma, K., and Wang, Z. (2020, January 19\u201324). Perceptual Quality Assessment of Smartphone Photography. Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition (CVPR), Copenhagen, Denmark.","DOI":"10.1109\/CVPR42600.2020.00373"},{"key":"ref_38","doi-asserted-by":"crossref","unstructured":"Snyder, D., Garcia-Romero, D., Povey, D., and Khudanpur, S. (2017). Deep Neural Network Embeddings for Text-Independent Speaker Verification, Interspeech.","DOI":"10.21437\/Interspeech.2017-620"},{"key":"ref_39","unstructured":"Paszke, A., Gross, S., Chintala, S., Chanan, G., Yang, E., DeVito, Z., Lin, Z., Desmaison, A., Antiga, L., and Lerer, A. (2017, January 4\u20139). Automatic Differentiation in Pytorch. Proceedings of the 31st Conference on Neural Information Processing Systems (NIPS 2017), Long Beach, CA, USA."},{"key":"ref_40","first-page":"2825","article-title":"Scikit-learn: Machine Learning in Python","volume":"12","author":"Pedregosa","year":"2011","journal-title":"J. Mach. Learn. Res."},{"key":"ref_41","doi-asserted-by":"crossref","unstructured":"He, K., Zhang, X., Ren, S., and Sun, J. (2015, January 7\u201313). Delving deep into rectifiers: Surpassing human-level performance on imagenet classification. Proceedings of the IEEE International Conference on Computer Vision (ICCV), Santiago, Chile.","DOI":"10.1109\/ICCV.2015.123"},{"key":"ref_42","doi-asserted-by":"crossref","first-page":"148","DOI":"10.1109\/JPROC.2015.2494218","article-title":"Taking the Human Out of the Loop: A Review of Bayesian Optimization","volume":"104","author":"Shahriari","year":"2016","journal-title":"Proc. IEEE"},{"key":"ref_43","doi-asserted-by":"crossref","first-page":"64","DOI":"10.1145\/2812802","article-title":"YFCC100M: The new data in multimedia research","volume":"59","author":"Thomee","year":"2016","journal-title":"Commun. ACM"},{"key":"ref_44","unstructured":"Ye, P., Kumar, J., Kang, L., and Doermann, D. (2012, January 16\u201321). Unsupervised feature learning framework for no-reference image quality assessment. Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition, Providence, RI, USA."}],"container-title":["Journal of Imaging"],"original-title":[],"language":"en","link":[{"URL":"https:\/\/www.mdpi.com\/2313-433X\/7\/3\/55\/pdf","content-type":"unspecified","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2025,10,11]],"date-time":"2025-10-11T05:35:20Z","timestamp":1760160920000},"score":1,"resource":{"primary":{"URL":"https:\/\/www.mdpi.com\/2313-433X\/7\/3\/55"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2021,3,13]]},"references-count":44,"journal-issue":{"issue":"3","published-online":{"date-parts":[[2021,3]]}},"alternative-id":["jimaging7030055"],"URL":"https:\/\/doi.org\/10.3390\/jimaging7030055","relation":{},"ISSN":["2313-433X"],"issn-type":[{"value":"2313-433X","type":"electronic"}],"subject":[],"published":{"date-parts":[[2021,3,13]]}}}