{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2026,4,24]],"date-time":"2026-04-24T23:21:08Z","timestamp":1777072868515,"version":"3.51.4"},"reference-count":56,"publisher":"MDPI AG","issue":"2","license":[{"start":{"date-parts":[[2023,2,7]],"date-time":"2023-02-07T00:00:00Z","timestamp":1675728000000},"content-version":"vor","delay-in-days":0,"URL":"https:\/\/creativecommons.org\/licenses\/by\/4.0\/"}],"content-domain":{"domain":[],"crossmark-restriction":false},"short-container-title":["Algorithms"],"abstract":"<jats:p>The problem solved in the article is connected with the increase in the efficiency of phraseological radio exchange message recognition, which sometimes takes place in conditions of increased tension for the pilot. For high-quality recognition, signal preprocessing methods are needed. The article considers new data preprocessing algorithms used to extract features from a speech message. In this case, two approaches were proposed. The first approach is building autocorrelation functions of messages based on the Fourier transform, the second one uses the idea of building autocorrelation portraits of speech signals. The proposed approaches are quite simple to implement, although they require cyclic operators, since they work with pairs of samples from the original signal. Approbation of the developed method was carried out with the problem of recognizing phraseological radio exchange messages in Russian. The algorithm with preliminary feature extraction provides a gain of 1.7% in recognition accuracy. The use of convolutional neural networks also provides an increase in recognition efficiency. The gain for autocorrelation portraits processing is about 3\u20134%. Quantization is used to optimize the proposed models. The algorithm\u2019s performance increased by 2.8 times after the quantization. It was also possible to increase accuracy of recognition by 1\u20132% using digital signal processing algorithms. An important feature of the proposed algorithms is the possibility of generalizing them to arbitrary data with time correlation. The speech message preprocessing algorithms discussed in this article are based on classical digital signal processing algorithms. The idea of constructing autocorrelation portraits based on the time series of a signal has a novelty. At the same time, this approach ensures high recognition accuracy. However, the study also showed that all the algorithms under consideration perform quite poorly under the influence of strong noise.<\/jats:p>","DOI":"10.3390\/a16020090","type":"journal-article","created":{"date-parts":[[2023,2,8]],"date-time":"2023-02-08T05:37:31Z","timestamp":1675834651000},"page":"90","update-policy":"https:\/\/doi.org\/10.3390\/mdpi_crossmark_policy","source":"Crossref","is-referenced-by-count":5,"title":["The Use of Correlation Features in the Problem of Speech Recognition"],"prefix":"10.3390","volume":"16","author":[{"ORCID":"https:\/\/orcid.org\/0000-0003-0735-7697","authenticated-orcid":false,"given":"Nikita","family":"Andriyanov","sequence":"first","affiliation":[{"name":"Data Analysis and Machine Learning Department, Financial University under the Government of the Russian Federation, pr-kt Leningradsky, 49\/2, 125167 Moscow, Russia"}],"role":[{"role":"author","vocabulary":"crossref"}]}],"member":"1968","published-online":{"date-parts":[[2023,2,7]]},"reference":[{"key":"ref_1","doi-asserted-by":"crossref","unstructured":"Parekh, D., Poddar, N., Rajpurkar, A., Chahal, M., Kumar, N., Joshi, G.P., and Cho, W. (2022). A Review on Autonomous Vehicles: Progress, Methods and Challenges. Electronics, 11.","DOI":"10.3390\/electronics11142162"},{"key":"ref_2","doi-asserted-by":"crossref","unstructured":"Khanum, A., Lee, C.-Y., and Yang, C.-S. (2022). Deep-Learning-Based Network for Lane Following in Autonomous Vehicles. Electronics, 11.","DOI":"10.3390\/electronics11193084"},{"key":"ref_3","doi-asserted-by":"crossref","unstructured":"Brunelli, M., Ditta, C.C., and Postorino, M.N. (2022). A Framework to Develop Urban Aerial Networks by Using a Digital Twin Approach. Drones, 6.","DOI":"10.3390\/drones6120387"},{"key":"ref_4","doi-asserted-by":"crossref","first-page":"1014","DOI":"10.1007\/978-3-030-29513-4_74","article-title":"Using Local Objects to Improve Estimation of Mobile Object Coordinates and Smoothing Trajectory of Movement by Autoregression with Multiple Roots","volume":"1038","author":"Andriyanov","year":"2020","journal-title":"Adv. Intell. Syst. Comput."},{"key":"ref_5","doi-asserted-by":"crossref","unstructured":"Jarray, R., Bouall\u00e8gue, S., Rezk, H., and Al-Dhaifallah, M. (2022). Parallel Multiobjective Multiverse Optimizer for Path Planning of Unmanned Aerial Vehicles in a Dynamic Environment with Moving Obstacles. Drones, 6.","DOI":"10.3390\/drones6120385"},{"key":"ref_6","doi-asserted-by":"crossref","first-page":"489","DOI":"10.1134\/S1054661822030026","article-title":"Combining Text and Image Analysis Methods for Solving Multimodal Classification Problems","volume":"32","author":"Andriyanov","year":"2022","journal-title":"Pattern Recognit. Image Anal."},{"key":"ref_7","doi-asserted-by":"crossref","unstructured":"Mukhamadiyev, A., Khujayarov, I., Djuraev, O., and Cho, J. (2022). Automatic Speech Recognition Method Based on Deep Learning Approaches for Uzbek Language. Sensors, 22.","DOI":"10.3390\/s22103683"},{"key":"ref_8","doi-asserted-by":"crossref","unstructured":"Ramos-P\u00e9rez, E., Alonso-Gonz\u00e1lez, P.J., and N\u00fa\u00f1ez-Vel\u00e1zquez, J.J. (2021). Multi-Transformer: A New Neural Network-Based Architecture for Forecasting S & P Volatility. Mathematics, 9.","DOI":"10.3390\/math9151794"},{"key":"ref_9","doi-asserted-by":"crossref","unstructured":"Andriyanov, N., and Papakostas, G. (2022, January 23\u201327). Optimization and Benchmarking of Convolutional Networks with Quantization and OpenVINO in Baggage Image Recognition. Proceedings of the 2022 VIII International Conference on Information Technology and Nanotechnology (ITNT), Samara, Russia.","DOI":"10.1109\/ITNT55410.2022.9848757"},{"key":"ref_10","doi-asserted-by":"crossref","unstructured":"Wu, X., Jin, Y., Wang, J., Qian, Q., and Guo, Y. (2022). MKD: Mixup-Based Knowledge Distillation for Mandarin End-to-End Speech Recognition. Algorithms, 15.","DOI":"10.3390\/a15050160"},{"key":"ref_11","doi-asserted-by":"crossref","unstructured":"Andriyanov, N., Dementiev, V., and Gladkikh, A. (2021, January 13\u201314). Analysis of the Pattern Recognition Efficiency on Non-Optical Images. Proceedings of the 2021 Ural Symposium on Biomedical Engineering, Radioelectronics and Information Technology (USBEREIT), Yekaterinburg, Russia.","DOI":"10.1109\/USBEREIT51232.2021.9455097"},{"key":"ref_12","doi-asserted-by":"crossref","unstructured":"Riz\u00e0 Porta, R., Sterchi, Y., and Schwaninger, A. (2022). How Realistic Is Threat Image Projection for X-ray Baggage Screening?. Sensors, 22.","DOI":"10.3390\/s22062220"},{"key":"ref_13","doi-asserted-by":"crossref","unstructured":"Ribas, D., Miguel, A., Ortega, A., and Lleida, E. (2022). Wiener Filter and Deep Neural Networks: A Well-Balanced Pair for Speech Enhancement. Appl. Sci., 12.","DOI":"10.3390\/app12189000"},{"key":"ref_14","doi-asserted-by":"crossref","unstructured":"Antonetti, A.E.d.S., Siqueira, L.T.D., Gobbo, M.P.d.A., Brasolotto, A.G., and Silverio, K.C.A. (2020). Relationship of Cepstral Peak Prominence-Smoothed and Long-Term Average Spectrum with Auditory\u2013Perceptual Analysis. Appl. Sci., 10.","DOI":"10.3390\/app10238598"},{"key":"ref_15","doi-asserted-by":"crossref","unstructured":"Andriyanov, N., and Andriyanov, D. (2021, January 13\u201315). Intelligent Processing of Voice Messages in Civil Aviation: Message Recognition and the Emotional State of the Speaker Analysis. Proceedings of the 2021 International Siberian Conference on Control and Communications (SIBCON), Kazan, Russia.","DOI":"10.1109\/SIBCON50419.2021.9438881"},{"key":"ref_16","first-page":"91","article-title":"Recognition of radio exchange voice messages in aviation based on correlation analysis","volume":"23","author":"Andriyanov","year":"2021","journal-title":"Izv. Samara Sci. Cent. Russ. Acad. Sci."},{"key":"ref_17","doi-asserted-by":"crossref","first-page":"436","DOI":"10.1038\/nature14539","article-title":"Deep learning","volume":"521","author":"LeCun","year":"2015","journal-title":"Nature"},{"key":"ref_18","doi-asserted-by":"crossref","unstructured":"Dhouib, A., Othman, A., El Ghoul, O., Khribi, M.K., and Al Sinani, A. (2022). Arabic Automatic Speech Recognition: A Systematic Literature Review. Appl. Sci., 12.","DOI":"10.3390\/app12178898"},{"key":"ref_19","doi-asserted-by":"crossref","unstructured":"Nallasamy, U., Metze, F., and Schultz, T. (2012, January 2\u20135). Active Learning for Accent Adaptation in Automatic Speech Recognition. Proceedings of the 2012 IEEE Spoken Language Technology Workshop (SLT), Miami, FL, USA.","DOI":"10.1109\/SLT.2012.6424250"},{"key":"ref_20","doi-asserted-by":"crossref","unstructured":"Wahyuni, E.S. (2017, January 1\u20132). Arabic Speech Recognition Using MFCC Feature Extraction and ANN Classification. Proceedings of the 2017 2nd International Conferences on Information Technology, Information Systems and Electrical Engineering (ICITISEE), Yogyakarta, Indonesia.","DOI":"10.1109\/ICITISEE.2017.8285499"},{"key":"ref_21","doi-asserted-by":"crossref","unstructured":"Trinh Van, L., Dao Thi Le, T., Le Xuan, T., and Castelli, E. (2022). Emotional Speech Recognition Using Deep Neural Networks. Sensors, 22.","DOI":"10.3390\/s22041414"},{"key":"ref_22","doi-asserted-by":"crossref","unstructured":"Satt, A., Rozenberg, S., and Hoory, R. (2017, January 20\u201324). Efficient Emotion Recognition from Speech Using Deep Learning on Spectrograms. Proceedings of the International Speech Communication Association (INTERSPEECH), Stockholm, Sweden.","DOI":"10.21437\/Interspeech.2017-200"},{"key":"ref_23","first-page":"1","article-title":"Testing of the Speech Recognition Systems Using Russian Language Models","volume":"2298","author":"Aksyonov","year":"2018","journal-title":"CEUR Workshop Proc."},{"key":"ref_24","doi-asserted-by":"crossref","unstructured":"Vazhenina, D., Kipyatkova, I., Markov, K., and Karpov, A. (2012, January 8\u201313). State-of-the-art speech recognition technologies for Russian language. HCCE\u201912. Proceedings of the 2012 Joint International Conference on Human-Centered Computer Environments, Aizu-Wakamatsu, Japan.","DOI":"10.1145\/2160749.2160763"},{"key":"ref_25","unstructured":"Bagley, S., Antonov, A., Meshkov, B., and Sukhanov, A. (2009, January 27\u201331). Statistical Distribution of Words in a Russian Text Collection. Proceedings of the Dialogue 2009, Bekasovo, Serbia."},{"key":"ref_26","doi-asserted-by":"crossref","unstructured":"Alqadasi, A.M.A., Sunar, M.S., Turaev, S., Abdulghafor, R., Hj Salam, M.S., Alashbi, A.A.S., Salem, A.A., and Ali, M.A.H. (2023). Rule-Based Embedded HMMs Phoneme Classification to Improve Qur\u2019anic Recitation Recognition. Electronics, 12.","DOI":"10.3390\/electronics12010176"},{"key":"ref_27","doi-asserted-by":"crossref","unstructured":"Oh, D., Park, J.-S., Kim, J.-H., and Jang, G.-J. (2021). Hierarchical Phoneme Classification for Improved Speech Recognition. Appl. Sci., 11.","DOI":"10.3390\/app11010428"},{"key":"ref_28","doi-asserted-by":"crossref","unstructured":"Liu, Z., Huang, Z., Wang, L., and Zhang, P. (2021). A Pronunciation Prior Assisted Vowel Reduction Detection Framework with Multi-Stream Attention Method. Appl. Sci., 11.","DOI":"10.3390\/app11188321"},{"key":"ref_29","doi-asserted-by":"crossref","unstructured":"Jeon, S., and Kim, M.S. (2022). Noise-Robust Multimodal Audio-Visual Speech Recognition System for Speech-Based Interaction Applications. Sensors, 22.","DOI":"10.3390\/s22207738"},{"key":"ref_30","doi-asserted-by":"crossref","unstructured":"Vazhenina, D., and Markov, K. (2020). End-to-End Noisy Speech Recognition Using Fourier and Hilbert Spectrum Features. Electronics, 9.","DOI":"10.3390\/electronics9071157"},{"key":"ref_31","doi-asserted-by":"crossref","unstructured":"Pervaiz, A., Hussain, F., Israr, H., Tahir, M.A., Raja, F.R., Baloch, N.K., Ishmanov, F., and Zikria, Y.B. (2020). Incorporating Noise Robustness in Speech Command Recognition by Noise Augmentation of Training Data. Sensors, 20.","DOI":"10.3390\/s20082326"},{"key":"ref_32","doi-asserted-by":"crossref","first-page":"012018","DOI":"10.1088\/1742-6596\/1661\/1\/012018","article-title":"The using of data augmentation in machine learning in image processing tasks in the face of data scarcity","volume":"1661","author":"Andriyanov","year":"2020","journal-title":"J. Phys. Conf. Ser."},{"key":"ref_33","doi-asserted-by":"crossref","unstructured":"Box, G., Jenkins, G., and Reinsel, G. (2008). Time Series Analysis, John Wiley & Sons, Inc.","DOI":"10.1002\/9781118619193"},{"key":"ref_34","unstructured":"Draper, N.R., and Smith, H. (1966). Applied Regression Analysis, Wiley."},{"key":"ref_35","first-page":"572173","article-title":"Autoregressive Prediction with Rolling Mechanism for Time Series Forecasting with Small Sample Size","volume":"2014","author":"Zhihua","year":"2014","journal-title":"Math. Probl. Eng."},{"key":"ref_36","doi-asserted-by":"crossref","unstructured":"Orzechowski, A., and Bombol, M. (2022). Energy Security, Sustainable Development and the Green Bond Market. Energies, 15.","DOI":"10.3390\/en15176218"},{"key":"ref_37","first-page":"1","article-title":"Time series Forecasting using Holt-Winters Exponential Smoothing","volume":"13","author":"Prajakta","year":"2004","journal-title":"Kanwal Rekhi Sch. Inf. Technol. J."},{"key":"ref_38","doi-asserted-by":"crossref","first-page":"464","DOI":"10.3390\/geomatics1040027","article-title":"Measuring Similarity of Deforestation Patterns in Time and Space across Differences in Resolution","volume":"1","author":"Suyamto","year":"2021","journal-title":"Geomatics"},{"key":"ref_39","first-page":"5681308","article-title":"Forecasting Drought Using Multilayer Perceptron Artificial Neural Network Model","volume":"2017","author":"Zulifqar","year":"2017","journal-title":"Adv. Meteorol."},{"key":"ref_40","doi-asserted-by":"crossref","first-page":"132306","DOI":"10.1016\/j.physd.2019.132306","article-title":"Fundamentals of Recurrent Neural Network (RNN) and Long Short-Term Memory (LSTM) Network","volume":"404","author":"Sherstinsky","year":"2020","journal-title":"Phys. D Nonlinear Phenom."},{"key":"ref_41","doi-asserted-by":"crossref","first-page":"139","DOI":"10.18287\/2412-6179-CO-922","article-title":"Detection of objects in the images: From likelihood relationships towards scalable and efficient neural networks","volume":"46","author":"Andriyanov","year":"2022","journal-title":"Comput. Opt."},{"key":"ref_42","doi-asserted-by":"crossref","unstructured":"Dua, S., Kumar, S.S., Albagory, Y., Ramalingam, R., Dumka, A., Singh, R., Rashid, M., Gehlot, A., Alshamrani, S.S., and AlGhamdi, A.S. (2022). Developing a Speech Recognition System for Recognizing Tonal Speech Signals Using a Convolutional Neural Network. Appl. Sci., 12.","DOI":"10.3390\/app12126223"},{"key":"ref_43","doi-asserted-by":"crossref","unstructured":"Salas-P\u00e1ez, C., Quintana-Romero, L., Mendoza-Gonz\u00e1lez, M.A., and \u00c1lvarez-Garc\u00eda, J. (2022). Analysis of Job Transitions in Mexico with Markov Chains in Discrete Time. Mathematics, 10.","DOI":"10.3390\/math10101693"},{"key":"ref_44","unstructured":"Yohannes, Y., and Webb, P. (1999). Classification and Regression Trees, CART: A User Manual for Identifying Indicators of Vulnerability to Famine and Chronic Food Insecurity, International Food Policy Research Institute."},{"key":"ref_45","first-page":"23","article-title":"Time series forecasting via genetic algorithm for turkish air transport market","volume":"9","author":"Pehlivanoglu","year":"2016","journal-title":"J. Aeronaut. Space Technol."},{"key":"ref_46","unstructured":"Wenzel, F., Galy-Fajou, T., Deutsch, M., and Kloft, M. (2017). Machine Learning and Knowledge Discovery in Databases: European Conference, ECML PKDD 2017, Skopje, Macedonia, 18\u201322 September 2017, Proceedings, Part I, Springer."},{"key":"ref_47","first-page":"10","article-title":"Algorithm based on the transfer function model and one-class classification for detecting the anomalous state of dams","volume":"6","author":"Kozionova","year":"2015","journal-title":"Inf. Control. Syst."},{"key":"ref_48","first-page":"246","article-title":"Identification anomalies the time series of metrics of project based on entropy measures","volume":"1","author":"Timina","year":"2017","journal-title":"Interact. Syst. Probl. Hum. Comput. Interact."},{"key":"ref_49","doi-asserted-by":"crossref","first-page":"245","DOI":"10.1109\/TPAMI.1987.4767898","article-title":"Image Estimation Using Doubly Stochastic Gaussian Random Field Models","volume":"9","author":"Woods","year":"1987","journal-title":"Pattern Anal. Mach. Intell."},{"key":"ref_50","doi-asserted-by":"crossref","first-page":"012188","DOI":"10.1088\/1742-6596\/1096\/1\/012188","article-title":"Ensuring the effectiveness of the taxi order service by mathematical modeling and machine learning","volume":"1096","author":"Danilov","year":"2018","journal-title":"J. Phys. Conf. Ser."},{"key":"ref_51","doi-asserted-by":"crossref","first-page":"183","DOI":"10.1007\/978-981-19-3444-5_16","article-title":"Development and Research of Intellectual Algorithms in Taxi Service Data Processing Based on Machine Learning and Modified K-means Method","volume":"Volume 309","author":"Andriyanov","year":"2022","journal-title":"Intelligent Decision Technologies. Smart Innovation, Systems and Technologies"},{"key":"ref_52","unstructured":"Armer, A.I. (2006). Modeling and Recognition of Speech Signals Against the Background of Intense Interference. [Ph.D. Thesis, Ulyanovsk State Technical University]."},{"key":"ref_53","unstructured":"Krasheninnikov, V.R., Lebedeva, E.Y., and Kapyrin, V.K. (2013, January 20\u201321). Variation of the boundaries of speech commands to improve the recognition of speech commands by their cross-correlation portraits. Proceedings of the Samara Scientific Center of the Russian Academy of Sciences, Samara, Russia."},{"key":"ref_54","first-page":"5511","article-title":"Automatic Speaker Recognition Using Mel-Frequency Cepstral Coefficients Through Machine Learning","volume":"71","author":"Ayvaz","year":"2022","journal-title":"Comput. Mater. Contin."},{"key":"ref_55","doi-asserted-by":"crossref","unstructured":"Khan, F., Tarimer, I., Alwageed, H.S., Karada\u011f, B.C., Fayaz, M., Abdusalomov, A.B., and Cho, Y.-I. (2022). Effect of Feature Selection on the Accuracy of Music Popularity Classification Using Machine Learning Algorithms. Electronics, 11.","DOI":"10.3390\/electronics11213518"},{"key":"ref_56","unstructured":"(2023, January 11). Audacity. Available online: https:\/\/www.audacityteam.org\/."}],"container-title":["Algorithms"],"original-title":[],"language":"en","link":[{"URL":"https:\/\/www.mdpi.com\/1999-4893\/16\/2\/90\/pdf","content-type":"unspecified","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2025,10,10]],"date-time":"2025-10-10T18:27:03Z","timestamp":1760120823000},"score":1,"resource":{"primary":{"URL":"https:\/\/www.mdpi.com\/1999-4893\/16\/2\/90"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2023,2,7]]},"references-count":56,"journal-issue":{"issue":"2","published-online":{"date-parts":[[2023,2]]}},"alternative-id":["a16020090"],"URL":"https:\/\/doi.org\/10.3390\/a16020090","relation":{},"ISSN":["1999-4893"],"issn-type":[{"value":"1999-4893","type":"electronic"}],"subject":[],"published":{"date-parts":[[2023,2,7]]}}}