{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2026,4,15]],"date-time":"2026-04-15T21:43:00Z","timestamp":1776289380536,"version":"3.50.1"},"reference-count":64,"publisher":"MDPI AG","issue":"3","license":[{"start":{"date-parts":[[2022,2,27]],"date-time":"2022-02-27T00:00:00Z","timestamp":1645920000000},"content-version":"vor","delay-in-days":0,"URL":"https:\/\/creativecommons.org\/licenses\/by\/4.0\/"}],"content-domain":{"domain":[],"crossmark-restriction":false},"short-container-title":["Future Internet"],"abstract":"<jats:p>Media authentication relies on the detection of inconsistencies that may indicate malicious editing in audio and video files. Traditionally, authentication processes are performed by forensics professionals using dedicated tools. There is rich research on the automation of this procedure, but the results do not yet guarantee the feasibility of providing automated tools. In the current approach, a computer-supported toolbox is presented, providing online functionality for assisting technically inexperienced users (journalists or the public) to investigate visually the consistency of audio streams. Several algorithms based on previous research have been incorporated on the backend of the proposed system, including a novel CNN model that performs a Signal-to-Reverberation-Ratio (SRR) estimation with a mean square error of 2.9%. The user can access the web application online through a web browser. After providing an audio\/video file or a YouTube link, the application returns as output a set of interactive visualizations that can allow the user to investigate the authenticity of the file. The visualizations are generated based on the outcomes of Digital Signal Processing and Machine Learning models. The files are stored in a database, along with their analysis results and annotation. Following a crowdsourcing methodology, users are allowed to contribute by annotating files from the dataset concerning their authenticity. The evaluation version of the web application is publicly available online.<\/jats:p>","DOI":"10.3390\/fi14030075","type":"journal-article","created":{"date-parts":[[2022,2,27]],"date-time":"2022-02-27T20:46:17Z","timestamp":1645994777000},"page":"75","update-policy":"https:\/\/doi.org\/10.3390\/mdpi_crossmark_policy","source":"Crossref","is-referenced-by-count":8,"title":["A Prototype Web Application to Support Human-Centered Audiovisual Content Authentication and Crowdsourcing"],"prefix":"10.3390","volume":"14","author":[{"ORCID":"https:\/\/orcid.org\/0000-0002-4801-0115","authenticated-orcid":false,"given":"Nikolaos","family":"Vryzas","sequence":"first","affiliation":[{"name":"Multidisciplinary Media & Mediated Communication Research Group (M3C), Aristotle University of Thessaloniki, 54124 Thessaloniki, Greece"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Anastasia","family":"Katsaounidou","sequence":"additional","affiliation":[{"name":"Multidisciplinary Media & Mediated Communication Research Group (M3C), Aristotle University of Thessaloniki, 54124 Thessaloniki, Greece"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"ORCID":"https:\/\/orcid.org\/0000-0003-2900-4657","authenticated-orcid":false,"given":"Lazaros","family":"Vrysis","sequence":"additional","affiliation":[{"name":"Multidisciplinary Media & Mediated Communication Research Group (M3C), Aristotle University of Thessaloniki, 54124 Thessaloniki, Greece"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Rigas","family":"Kotsakis","sequence":"additional","affiliation":[{"name":"Multidisciplinary Media & Mediated Communication Research Group (M3C), Aristotle University of Thessaloniki, 54124 Thessaloniki, Greece"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"ORCID":"https:\/\/orcid.org\/0000-0001-7923-9361","authenticated-orcid":false,"given":"Charalampos","family":"Dimoulas","sequence":"additional","affiliation":[{"name":"Multidisciplinary Media & Mediated Communication Research Group (M3C), Aristotle University of Thessaloniki, 54124 Thessaloniki, Greece"}],"role":[{"role":"author","vocabulary":"crossref"}]}],"member":"1968","published-online":{"date-parts":[[2022,2,27]]},"reference":[{"key":"ref_1","doi-asserted-by":"crossref","unstructured":"Katsaounidou, A.N., and Dimoulas, C.A. (2018). Integrating Content Authentication Support in Media Services. Encyclopedia of Information Science and Technology, IGI Global. [4th ed.].","DOI":"10.4018\/978-1-5225-2255-3.ch254"},{"key":"ref_2","unstructured":"Katsaounidou, A., and Dimoulas, C. (2018, January 18\u201319). The Role of media educator on the age of misinformation Crisis. Proceedings of the EJTA Teachers\u2019 Conference on Crisis Reporting, Thessaloniki, Greece."},{"key":"ref_3","doi-asserted-by":"crossref","unstructured":"Katsaounidou, A., Dimoulas, C., and Veglis, A. (2019). Cross-Media Authentication and Verification: Emerging Research and Opportunities, IGI Global.","DOI":"10.4018\/978-1-5225-5592-6"},{"key":"ref_4","doi-asserted-by":"crossref","unstructured":"Katsaounidou, A., Vrysis, L., Kotsakis, R., Dimoulas, C., and Veglis, A. (2019). MAthE the game: A serious game for education and training in news verification. Educ. Sci., 9.","DOI":"10.3390\/educsci9020155"},{"key":"ref_5","doi-asserted-by":"crossref","unstructured":"Katsaounidou, A., Vryzas, N., Kotsakis, R., and Dimoulas, C. (2019). Multimodal News authentication as a service: The \u201cTrue News\u201d Extension. J. Educ. Innov. Commun., 11\u201326.","DOI":"10.34097\/jeicom_SI_Dec2019-1"},{"key":"ref_6","doi-asserted-by":"crossref","first-page":"e05808","DOI":"10.1016\/j.heliyon.2020.e05808","article-title":"News authentication and tampered images: Evaluating the photo-truth impact through image verification algorithms","volume":"6","author":"Katsaounidou","year":"2020","journal-title":"Heliyon"},{"key":"ref_7","unstructured":"Vryzas, N., Katsaounidou, A., Kotsakis, R., Dimoulas, C.A., and Kalliris, G. (2018, January 23\u201326). Investigation of audio tampering in broadcast content. Proceedings of the Audio Engineering Society Convention 144, Milan, Italy."},{"key":"ref_8","unstructured":"Vryzas, N., Katsaounidou, A., Kotsakis, R., Dimoulas, C.A., and Kalliris, G. (2019, January 20\u201323). Audio-driven multimedia content authentication as a service. Proceedings of the Audio Engineering Society Convention 146, Dublin, Ireland."},{"key":"ref_9","unstructured":"Bakker, P. (2011, January 8\u20139). New journalism 3.0\u2014Aggregation, content farms, and Huffinization: The rise of low-pay and no-pay journalism. Proceedings of the Future of Journalism Conference, Cardiff, UK."},{"key":"ref_10","unstructured":"Graves, L., and Cherubini, F. (2016). The Rise of Fact-Checking Sites in Europe, Reuters Institute for the Study of Journalism."},{"key":"ref_11","doi-asserted-by":"crossref","first-page":"154","DOI":"10.1080\/21670811.2017.1345645","article-title":"Fake news and the economy of emotions: Problems, causes, solutions","volume":"6","author":"Bakir","year":"2018","journal-title":"Digit. Journal."},{"key":"ref_12","first-page":"41","article-title":"Big data analytics: Challenges and applications for text, audio, video, and social media data\u201d","volume":"5","author":"Verma","year":"2016","journal-title":"Int. J. Soft Comput. Artif. Intell. Appl."},{"key":"ref_13","doi-asserted-by":"crossref","unstructured":"Vlachos, A., and Riedel, S. (2014, January 26). Fact checking: Task definition and dataset construction. Proceedings of the ACL 2014 Workshop on Language Technologies and Computational Social Science, Baltimore, MD, USA.","DOI":"10.3115\/v1\/W14-2508"},{"key":"ref_14","doi-asserted-by":"crossref","first-page":"4801","DOI":"10.1007\/s11042-016-3795-2","article-title":"Large-scale evaluation of splicing localization algorithms for web images","volume":"76","author":"Zampoglou","year":"2017","journal-title":"Multimed. Tools Appl."},{"key":"ref_15","doi-asserted-by":"crossref","unstructured":"Zampoglou, M., Papadopoulos, S., and Kompatsiaris, Y. (July, January 29). Detecting image splicing in the wild (web). Proceedings of the 2015 IEEE International Conference on Multimedia & Expo Workshops, Turin, Italy.","DOI":"10.1109\/ICMEW.2015.7169839"},{"key":"ref_16","doi-asserted-by":"crossref","first-page":"8","DOI":"10.1016\/j.diin.2016.06.003","article-title":"Digital video tampering detection: An overview of passive techniques","volume":"18","author":"Sitara","year":"2016","journal-title":"Digit. Investig."},{"key":"ref_17","doi-asserted-by":"crossref","unstructured":"Grigoras, C., and Smith, J.M. (2013). Audio Enhancement and Authentication. Encyclopedia of Forensic Sciences, Elsevier.","DOI":"10.1016\/B978-0-12-382165-2.00128-8"},{"key":"ref_18","doi-asserted-by":"crossref","first-page":"84","DOI":"10.1109\/MSP.2008.931080","article-title":"Audio forensic examination","volume":"26","author":"Maher","year":"2009","journal-title":"IEEE Signal Process. Mag."},{"key":"ref_19","first-page":"3","article-title":"Authentication of forensic audio recordings","volume":"38","author":"Koenig","year":"1990","journal-title":"J. Audio Eng. Soc."},{"key":"ref_20","doi-asserted-by":"crossref","first-page":"1009","DOI":"10.1007\/s11042-016-4277-2","article-title":"Digital multimedia audio forensics: Past, present and future","volume":"77","author":"Zakariah","year":"2018","journal-title":"Multimed. Tools Appl."},{"key":"ref_21","doi-asserted-by":"crossref","first-page":"50","DOI":"10.1109\/MMUL.2011.74","article-title":"Current developments and future trends in audio authentication","volume":"19","author":"Gupta","year":"2011","journal-title":"IEEE Multimed."},{"key":"ref_22","doi-asserted-by":"crossref","first-page":"534","DOI":"10.1109\/TIFS.2010.2051270","article-title":"Audio authenticity: Detecting ENF discontinuity with high precision phase analysis","volume":"5","author":"Biscainho","year":"2010","journal-title":"IEEE Trans. Inf. Forensics Secur."},{"key":"ref_23","first-page":"643","article-title":"Applications of ENF analysis in forensic authentication of digital audio and video recordings","volume":"57","author":"Grigoras","year":"2009","journal-title":"J. Audio Eng. Soc."},{"key":"ref_24","unstructured":"Brixen, E.B. (2007, January 5\u20138). Techniques for the authentication of digital audio recordings. Proceedings of the Audio Engineering Society Convention 122, Vienna, Austria."},{"key":"ref_25","doi-asserted-by":"crossref","first-page":"1003","DOI":"10.1109\/TIFS.2016.2516824","article-title":"Audio authentication by exploring the absolute-error-map of ENF signals","volume":"11","author":"Hua","year":"2016","journal-title":"IEEE Trans. Inf. Forensics Secur."},{"key":"ref_26","doi-asserted-by":"crossref","unstructured":"Malik, H., and Farid, H. (2010, January 14\u201319). Audio forensics from acoustic reverberation. Proceedings of the IEEE International Conference on Acoustics, Speech and Signal Processing, Dallas, TX, USA.","DOI":"10.1109\/ICASSP.2010.5495479"},{"key":"ref_27","doi-asserted-by":"crossref","first-page":"1746","DOI":"10.1109\/TIFS.2013.2278843","article-title":"Audio recording location identification using acoustic environment signature","volume":"8","author":"Zhao","year":"2013","journal-title":"IEEE Trans. Inf. Forensics Secur."},{"key":"ref_28","doi-asserted-by":"crossref","unstructured":"Buchholz, R., Kraetzer, C., and Dittmann, J. (2009, January 8\u201310). Microphone classification using Fourier coefficients. Proceedings of the International Workshop on Information Hiding, Darmstadt, Germany.","DOI":"10.1007\/978-3-642-04431-1_17"},{"key":"ref_29","doi-asserted-by":"crossref","unstructured":"Garcia-Romero, D., and Espy-Wilson, C.Y. (2010, January 14\u201319). Automatic acquisition device identification from speech recordings. Proceedings of the 2010 IEEE International Conference on Acoustics, Speech and Signal Processing, Dallas, TX, USA.","DOI":"10.1109\/ICASSP.2010.5495407"},{"key":"ref_30","unstructured":"Hafeez, A., Malik, H., and Mahmood, K. (2017, January 15\u201317). Performance of blind microphone recognition algorithms in the presence of anti-forensic attacks. Proceedings of the 2017 AES International Conference on Audio Forensics, Arlington, VA, USA."},{"key":"ref_31","doi-asserted-by":"crossref","first-page":"1827","DOI":"10.1109\/TIFS.2013.2280888","article-title":"Acoustic environment identification and its applications to audio forensics","volume":"8","author":"Malik","year":"2013","journal-title":"IEEE Trans. Inf. Forensics Secur."},{"key":"ref_32","doi-asserted-by":"crossref","unstructured":"Narkhede, M., and Patole, R. (2019). Acoustic scene identification for audio authentication. Soft Computing and Signal Processing, Springer.","DOI":"10.1007\/978-981-13-3600-3_56"},{"key":"ref_33","doi-asserted-by":"crossref","unstructured":"Patole, R.K., Rege, P.P., and Suryawanshi, P. (2016, January 19\u201321). Acoustic environment identification using blind de-reverberation. Proceedings of the 2016 International Conference on Computing, Analytics and Security Trends (CAST), Pune, India.","DOI":"10.1109\/CAST.2016.7915019"},{"key":"ref_34","doi-asserted-by":"crossref","first-page":"123","DOI":"10.1016\/j.ins.2012.10.013","article-title":"MP3 audio steganalysis","volume":"231","author":"Qiao","year":"2013","journal-title":"Inf. Sci."},{"key":"ref_35","first-page":"75410K","article-title":"Detecting double compression of audio signal","volume":"Volume 7541","author":"Yang","year":"2010","journal-title":"Media Forensics and Security II, Proceedings of the IS&T\/SPIE Electronic Imaging, San Jose, CA, USA, 17\u201321 January 2010"},{"key":"ref_36","doi-asserted-by":"crossref","first-page":"291","DOI":"10.1007\/s12559-010-9045-4","article-title":"Detection of double MP3 compression","volume":"2","author":"Liu","year":"2010","journal-title":"Cogn. Comput."},{"key":"ref_37","doi-asserted-by":"crossref","unstructured":"Seichter, D., Cuccovillo, L., and Aichroth, P. (2016, January 20\u201325). AAC encoding detection and bitrate estimation using a convolutional neural network. Proceedings of the 2016 IEEE International Conference on Acoustics, Speech and Signal Processing (ICASSP), Shanghai, China.","DOI":"10.1109\/ICASSP.2016.7472041"},{"key":"ref_38","unstructured":"Lacroix, J., Prime, Y., Remy, A., and Derrien, O. (November, January 29). Lossless audio checker: A software for the detection of upscaling, upsampling, and transcoding in lossless musical tracks. Proceedings of the Audio Engineering Society Convention 139, New York, NY, USA."},{"key":"ref_39","unstructured":"G\u00e4rtner, D., Dittmar, C., Aichroth, P., Cuccovillo, L., Mann, S., and Schuller, G. (2014, January 26\u201329). Efficient cross-codec framing grid analysis for audio tampering detection. Proceedings of the Audio Engineering Society Convention 136, Berlin, Germany."},{"key":"ref_40","doi-asserted-by":"crossref","unstructured":"Hennequin, R., Royo-Letelier, J., and Moussallam, M. (2017, January 5\u20139). Codec independent lossy audio compression detection. Proceedings of the 2017 IEEE International Conference on Acoustics, Speech and Signal Processing (ICASSP), New Orleans, LA, USA.","DOI":"10.1109\/ICASSP.2017.7952251"},{"key":"ref_41","doi-asserted-by":"crossref","first-page":"85","DOI":"10.1016\/j.dsp.2014.11.003","article-title":"Identification of AMR decompressed audio","volume":"37","author":"Luo","year":"2015","journal-title":"Digit. Signal Process."},{"key":"ref_42","doi-asserted-by":"crossref","first-page":"304","DOI":"10.22452\/mjcs.vol32no4.4","article-title":"Authentication of mp4 file by perceptual hash and data hiding","volume":"32","author":"Maung","year":"2019","journal-title":"Malays. J. Comput. Sci."},{"key":"ref_43","doi-asserted-by":"crossref","first-page":"525","DOI":"10.1080\/00450618.2017.1296186","article-title":"A novel audio forensic data-set for digital multimedia forensics","volume":"50","author":"Khan","year":"2018","journal-title":"Aust. J. Forensic Sci."},{"key":"ref_44","unstructured":"G\u00e4rtner, D., Cuccovillo, L., Mann, S., and Aichroth, P. (2014, January 12\u201314). A multi-codec audio dataset for codec analysis and tampering detection. Proceedings of the Audio Engineering Society Conference: 54th International Conference: Audio Forensics: Techniques, Technologies and Practice, London, UK."},{"key":"ref_45","doi-asserted-by":"crossref","first-page":"12843","DOI":"10.1109\/ACCESS.2017.2717842","article-title":"Blind detection of copy-move forgery in digital audio forensics","volume":"5","author":"Imran","year":"2017","journal-title":"IEEE Access"},{"key":"ref_46","unstructured":"Madry, A., Makelov, A., Schmidt, L., Tsipras, D., and Vladu, A. (2017). Towards deep learning models resistant to adversarial attacks. arXiv."},{"key":"ref_47","unstructured":"Vryzas, N. (2020). Audiovisual Stream Analysis and Management Automation in Digital Media and Mediated Communication. [Ph.D. Dissertation, Aristotle University of Thessaloniki]."},{"key":"ref_48","doi-asserted-by":"crossref","first-page":"66","DOI":"10.17743\/jaes.2019.0058","article-title":"1D\/2D Deep CNNs vs. Temporal Feature Integration for General Audio Classification","volume":"68","author":"Vrysis","year":"2020","journal-title":"J. Audio Eng. Soc."},{"key":"ref_49","doi-asserted-by":"crossref","unstructured":"Brabham, D.C. (2013). Crowdsourcing, MIT Press.","DOI":"10.7551\/mitpress\/9693.001.0001"},{"key":"ref_50","doi-asserted-by":"crossref","first-page":"227","DOI":"10.17743\/jaes.2021.0001","article-title":"Enhanced Temporal Feature Integration in Audio Semantics via Alpha-Stable Modeling","volume":"69","author":"Vrysis","year":"2021","journal-title":"J. Audio Eng. Soc."},{"key":"ref_51","doi-asserted-by":"crossref","first-page":"410","DOI":"10.3390\/acoustics1020023","article-title":"An enhanced temporal feature integration method for environmental sound recognition","volume":"1","author":"Bountourakis","year":"2019","journal-title":"Acoustics"},{"key":"ref_52","unstructured":"Vrysis, L., Thoidis, I., Dimoulas, C., and Papanikolaou, G. (2020, January 2\u20135). Experimenting with 1D CNN Architectures for Generic Audio Classification. Proceedings of the Audio Engineering Society Convention 148, Vienna, Austria."},{"key":"ref_53","unstructured":"Vrysis, L., Vryzas, N., Sidiropoulos, E., Avraam, E., and Dimoulas, C.A. (2019, January 20\u201323). jReporter: A Smart Voice-Recording Mobile Application. Proceedings of the Audio Engineering Society Convention 146, Dublin, Ireland."},{"key":"ref_54","doi-asserted-by":"crossref","first-page":"1072","DOI":"10.17743\/jaes.2018.0066","article-title":"Analysis of 2d feature spaces for deep learning-based speech recognition","volume":"66","author":"Korvel","year":"2018","journal-title":"J. Audio Eng. Soc."},{"key":"ref_55","doi-asserted-by":"crossref","unstructured":"Ciaburro, G. (2020). Sound event detection in underground parking garage using convolutional neural network. Big Data Cogn. Comput., 4.","DOI":"10.3390\/bdcc4030020"},{"key":"ref_56","doi-asserted-by":"crossref","unstructured":"Ciaburro, G., and Iannace, G. (2020). Improving smart cities safety using sound events detection based on deep neural network algorithms. Informatics, 7.","DOI":"10.3390\/informatics7030023"},{"key":"ref_57","doi-asserted-by":"crossref","first-page":"189","DOI":"10.1177\/0165551512437638","article-title":"Towards an integrated crowdsourcing definition","volume":"38","year":"2012","journal-title":"J. Inf. Sci."},{"key":"ref_58","doi-asserted-by":"crossref","first-page":"1042","DOI":"10.17743\/jaes.2016.0051","article-title":"Crowdsourcing audio semantics by means of hybrid bimodal segmentation with hierarchical classification","volume":"64","author":"Vrysis","year":"2016","journal-title":"J. Audio Eng. Soc."},{"key":"ref_59","doi-asserted-by":"crossref","unstructured":"Vrysis, L., Tsipas, N., Dimoulas, C., and Papanikolaou, G. (2015, January 7\u20139). Mobile audio intelligence: From real time segmentation to crowd sourced semantics. Proceedings of the Audio Mostly 2015 on Interaction with Sound, Thessaloniki, Greece.","DOI":"10.1145\/2814895.2814906"},{"key":"ref_60","doi-asserted-by":"crossref","unstructured":"Cartwright, M., Dove, G., M\u00e9ndez M\u00e9ndez, A.E., Bello, J.P., and Nov, O. (2019, January 4\u20139). Crowdsourcing multi-label audio annotation tasks with citizen scientists. Proceedings of the 2019 CHI Conference on Human Factors in Computing Systems, Glasgow, UK.","DOI":"10.1145\/3290605.3300522"},{"key":"ref_61","doi-asserted-by":"crossref","unstructured":"Vrysis, L., Vryzas, N., Kotsakis, R., Saridou, T., Matsiola, M., Veglis, A., Arcila-Calder\u00f3n, C., and Dimoulas, C. (2021). A Web Interface for Analyzing Hate Speech. Future Internet, 13.","DOI":"10.3390\/fi13030080"},{"key":"ref_62","unstructured":"Chollet, F., Eldeeb, A., Bursztein, E., Jin, H., Watson, M., and Zhu, Q.S. (2015). Keras, GitHub. Available online: https:\/\/github.com\/fchollet\/keras."},{"key":"ref_63","doi-asserted-by":"crossref","unstructured":"McFee, B., Raffel, C., Liang, D., Ellis, D.P., McVicar, M., Battenberg, E., and Nieto, O. (2015, January 6\u201312). librosa: Audio and music signal analysis in Python. Proceedings of the 14th Python in Science Conference, Austin, TX, USA.","DOI":"10.25080\/Majora-7b98e3ed-003"},{"key":"ref_64","unstructured":"Collette, A. (2013). Python and HDF5: Unlocking Scientific Data, O\u2019Reilly Media, Inc."}],"container-title":["Future Internet"],"original-title":[],"language":"en","link":[{"URL":"https:\/\/www.mdpi.com\/1999-5903\/14\/3\/75\/pdf","content-type":"unspecified","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2025,10,10]],"date-time":"2025-10-10T22:28:38Z","timestamp":1760135318000},"score":1,"resource":{"primary":{"URL":"https:\/\/www.mdpi.com\/1999-5903\/14\/3\/75"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2022,2,27]]},"references-count":64,"journal-issue":{"issue":"3","published-online":{"date-parts":[[2022,3]]}},"alternative-id":["fi14030075"],"URL":"https:\/\/doi.org\/10.3390\/fi14030075","relation":{},"ISSN":["1999-5903"],"issn-type":[{"value":"1999-5903","type":"electronic"}],"subject":[],"published":{"date-parts":[[2022,2,27]]}}}