{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2026,7,6]],"date-time":"2026-07-06T06:15:19Z","timestamp":1783318519855,"version":"3.54.6"},"reference-count":165,"publisher":"Springer Science and Business Media LLC","issue":"12","license":[{"start":{"date-parts":[[2026,6,1]],"date-time":"2026-06-01T00:00:00Z","timestamp":1780272000000},"content-version":"tdm","delay-in-days":0,"URL":"https:\/\/creativecommons.org\/licenses\/by\/4.0"},{"start":{"date-parts":[[2026,6,13]],"date-time":"2026-06-13T00:00:00Z","timestamp":1781308800000},"content-version":"vor","delay-in-days":12,"URL":"https:\/\/creativecommons.org\/licenses\/by\/4.0"}],"funder":[{"name":"Swinburne University of Technology"}],"content-domain":{"domain":["link.springer.com"],"crossmark-restriction":false},"short-container-title":["Neural Comput &amp; Applic"],"published-print":{"date-parts":[[2026,6]]},"abstract":"<jats:title>Abstract<\/jats:title>\n                  <jats:p>Speech Emotion Recognition (SER) is an advanced technology for developing intuitive and empathetic human-computer interfaces (HCI). While traditional SER systems have achievement a certain degree of succeed in recognising basic emotions from acted speech in a closed environment, real-world applications necessitate the recognition of more complex emotions. This paper presents a systematic review of deep learning approaches in SER from 2019 to the present, following the PRISMA guidelines, with a specific focus on the bridge between basic and complex SER within unimodal (audio-only) and multimodal frameworks. Analysis was done on the landscape of emotion models, datasets, and state-of-the-art (SOTA) model architectures, including CNNs, RNNs, Transformers, and their hybrids. The results reveal that deep learning has improved performance; the following hybrid models improved considerably; however, unimodal models still struggle with the subtle and often overlapping acoustic features of complex emotions. In contrast, multimodal models that leverage complementary information are consistently superior. Nevertheless, challenges remain, such as the over-reliance on a limited range of non-naturalistic datasets, the subjectivity associated with labelling complex emotions, and models not generalising to the variability in the real world. Finally, a conclusion is drawn by offering a strategic roadmap to guide the continuation of research in recognising complex emotions, including the efficient creation of naturalistic, large datasets for future modelling, the development of more advanced techniques for multimodal fusion, and the targeting of unconsidered but available acoustic features to enhance the modelling of the complexity of human emotions.<\/jats:p>","DOI":"10.1007\/s00521-026-12186-w","type":"journal-article","created":{"date-parts":[[2026,6,13]],"date-time":"2026-06-13T02:58:12Z","timestamp":1781319492000},"update-policy":"https:\/\/doi.org\/10.1007\/springer_crossmark_policy","source":"Crossref","is-referenced-by-count":0,"title":["Speech emotion recognition using deep learning: from basic to complex emotions in unimodal and multimodal frameworks"],"prefix":"10.1007","volume":"38","author":[{"given":"Rachel Si Ting","family":"Lai","sequence":"first","affiliation":[],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Lau Bee","family":"Theng","sequence":"additional","affiliation":[],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Mark Kit Tsun","family":"Tee","sequence":"additional","affiliation":[],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Colin Choon Lin","family":"Tan","sequence":"additional","affiliation":[],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Caslon","family":"Chua","sequence":"additional","affiliation":[],"role":[{"vocabulary":"crossref","role":"author"}]}],"member":"297","published-online":{"date-parts":[[2026,6,13]]},"reference":[{"key":"12186_CR1","doi-asserted-by":"publisher","first-page":"1526","DOI":"10.1111\/1460-6984.12879","volume":"58","author":"C Alighieri","year":"2023","unstructured":"Alighieri C, Bettens K, Verbeke J, Van Lierde K (2023) Sometimes I feel sad\u2019: A qualitative study on children\u2019s perceptions with cleft palate speech and language therapy. Intl J Lang Comm Disor 58:1526\u20131538. https:\/\/doi.org\/10.1111\/1460-6984.12879","journal-title":"Intl J Lang Comm Disor"},{"key":"12186_CR2","doi-asserted-by":"publisher","first-page":"487","DOI":"10.1044\/2021_JSLHR-21-00234","volume":"65","author":"C Kao","year":"2022","unstructured":"Kao C, Sera MD, Zhang Y (2022) Emotional Speech Processing in 3- to 12-Month-Old Infants: Influences of Emotion Categories and Acoustic Parameters. J Speech Lang Hear Res 65:487\u2013500. https:\/\/doi.org\/10.1044\/2021_JSLHR-21-00234","journal-title":"J Speech Lang Hear Res"},{"key":"12186_CR3","unstructured":"Don HH, Sandra E, Honckenbury (2007) Discovering Psychology. Worth"},{"key":"12186_CR4","doi-asserted-by":"crossref","unstructured":"Ekman P, Thank I, Davidson R et al (1992) Are There Basic Emotions?","DOI":"10.1037\/0033-295X.99.3.550"},{"key":"12186_CR5","doi-asserted-by":"publisher","unstructured":"Berrios R (2019) What is complex\/emotional about emotional complexity? Front Psychol 10. https:\/\/doi.org\/10.3389\/fpsyg.2019.01606","DOI":"10.3389\/fpsyg.2019.01606"},{"key":"12186_CR6","doi-asserted-by":"publisher","first-page":"496","DOI":"10.1016\/j.concog.2008.03.014","volume":"17","author":"A Zinck","year":"2008","unstructured":"Zinck A (2008) Self-referential emotions. Conscious Cogn 17:496\u2013505. https:\/\/doi.org\/10.1016\/j.concog.2008.03.014","journal-title":"Conscious Cogn"},{"key":"12186_CR7","doi-asserted-by":"publisher","first-page":"150","DOI":"10.1016\/j.ins.2019.09.005","volume":"509","author":"L Chen","year":"2020","unstructured":"Chen L, Su W, Feng Y et al (2020) Two-layer fuzzy multiple random forest for speech emotion recognition in human-robot interaction. Inf Sci 509:150\u2013163. https:\/\/doi.org\/10.1016\/j.ins.2019.09.005","journal-title":"Inf Sci"},{"key":"12186_CR8","doi-asserted-by":"publisher","unstructured":"Weninger F, Eyben F, Schuller BW et al (2013) On the acoustics of emotion in audio: What speech, music, and sound have in common. Front Psychol 4. https:\/\/doi.org\/10.3389\/fpsyg.2013.00292","DOI":"10.3389\/fpsyg.2013.00292"},{"key":"12186_CR9","doi-asserted-by":"publisher","first-page":"56","DOI":"10.1016\/j.specom.2019.12.001","volume":"116","author":"MB Ak\u00e7ay","year":"2020","unstructured":"Ak\u00e7ay MB, O\u011fuz K (2020) Speech emotion recognition: Emotional models, databases, features, preprocessing methods, supporting modalities, and classifiers. Speech Commun 116:56\u201376. https:\/\/doi.org\/10.1016\/j.specom.2019.12.001","journal-title":"Speech Commun"},{"key":"12186_CR10","doi-asserted-by":"publisher","DOI":"10.1109\/ACCESS.2024.3476960","author":"GH Mohmad","year":"2024","unstructured":"Mohmad GH, Delhibabu R (2024) Speech Databases, speech features and classifiers in speech emotion recognition: A Review. IEEE Access. https:\/\/doi.org\/10.1109\/ACCESS.2024.3476960","journal-title":"IEEE Access"},{"key":"12186_CR11","doi-asserted-by":"publisher","unstructured":"Madanian S, Chen T, Adeleye O et al (2023) Speech emotion recognition using machine learning \u2014 A systematic review. Intell Syst Appl 20. https:\/\/doi.org\/10.1016\/j.iswa.2023.200266","DOI":"10.1016\/j.iswa.2023.200266"},{"key":"12186_CR12","doi-asserted-by":"publisher","DOI":"10.1109\/RADIOELEK.2019.8733432","volume-title":"Deep Learning Techniques for Speech Emotion Recognition: A Review; Deep Learning Techniques for Speech","author":"SK Pandey","year":"2019","unstructured":"Pandey SK, Shekhawat HS, Prasanna SRM (2019) Deep Learning Techniques for Speech Emotion Recognition: A Review; Deep Learning Techniques for Speech. A Review, Emotion Recognition"},{"key":"12186_CR13","doi-asserted-by":"publisher","first-page":"29307","DOI":"10.1007\/s11042-023-14656-y","volume":"82","author":"K Kaur","year":"2023","unstructured":"Kaur K, Singh P (2023) Trends in speech emotion recognition: a comprehensive survey. Multimedia Tools Appl 82:29307\u201329351. https:\/\/doi.org\/10.1007\/s11042-023-14656-y","journal-title":"Multimedia Tools Appl"},{"key":"12186_CR14","doi-asserted-by":"publisher","first-page":"312","DOI":"10.1016\/j.bspc.2018.08.035","volume":"47","author":"J Zhao","year":"2019","unstructured":"Zhao J, Mao X, Chen L (2019) Speech emotion recognition using deep 1D & 2D CNN LSTM networks. Biomed Signal Process Control 47:312\u2013323. https:\/\/doi.org\/10.1016\/j.bspc.2018.08.035","journal-title":"Biomed Signal Process Control"},{"key":"12186_CR15","doi-asserted-by":"publisher","first-page":"36018","DOI":"10.1109\/ACCESS.2022.3163856","volume":"10","author":"F Andayani","year":"2022","unstructured":"Andayani F, Theng LB, Tsun MT, Chua C (2022) Hybrid LSTM-Transformer Model for Emotion Recognition From Speech Audio Files. IEEE Access 10:36018\u201336027. https:\/\/doi.org\/10.1109\/ACCESS.2022.3163856","journal-title":"IEEE Access"},{"key":"12186_CR16","doi-asserted-by":"publisher","first-page":"19999","DOI":"10.1109\/ACCESS.2021.3054345","volume":"9","author":"M Ezz-Eldin","year":"2021","unstructured":"Ezz-Eldin M, Khalaf AAM, Hamed HFA, Hussein AI (2021) Efficient Feature-Aware Hybrid Model of Deep Learning Architectures for Speech Emotion Recognition. IEEE Access 9:19999\u201320011. https:\/\/doi.org\/10.1109\/ACCESS.2021.3054345","journal-title":"IEEE Access"},{"key":"12186_CR17","doi-asserted-by":"publisher","first-page":"37603","DOI":"10.1007\/s11042-023-16849-x","volume":"83","author":"S Mishra","year":"2024","unstructured":"Mishra S, Bhatnagar N, Prakasam P, T. R S (2024) Speech emotion recognition and classification using hybrid deep CNN and BiLSTM model. Multimedia Tools Appl 83:37603\u201337620. https:\/\/doi.org\/10.1007\/s11042-023-16849-x","journal-title":"Multimedia Tools Appl"},{"key":"12186_CR18","doi-asserted-by":"publisher","first-page":"124","DOI":"10.1007\/s11760-024-03574-7","volume":"19","author":"J Ning","year":"2025","unstructured":"Ning J, Zhang W (2025) Speech-based emotion recognition using a hybrid RNN-CNN network. SIViP 19:124. https:\/\/doi.org\/10.1007\/s11760-024-03574-7","journal-title":"SIViP"},{"key":"12186_CR19","doi-asserted-by":"publisher","first-page":"47795","DOI":"10.1109\/ACCESS.2021.3068045","volume":"9","author":"TM Wani","year":"2021","unstructured":"Wani TM, Gunawan TS, Qadri SAA et al (2021) A Comprehensive Review of Speech Emotion Recognition Systems. IEEE Access 9:47795\u201347814. https:\/\/doi.org\/10.1109\/ACCESS.2021.3068045","journal-title":"IEEE Access"},{"key":"12186_CR20","doi-asserted-by":"publisher","unstructured":"George SM, Muhamed Ilyas P (2024) A review on speech emotion recognition: A survey, recent advances, challenges, and the influence of noise. Neurocomputing 568. https:\/\/doi.org\/10.1016\/j.neucom.2023.127015","DOI":"10.1016\/j.neucom.2023.127015"},{"key":"12186_CR21","first-page":"30","volume":"2","author":"M Anandappa","year":"2024","unstructured":"Anandappa M, Kavita Mudnal M (2024) Analysis Of Emotions Through Speech Recognition. J Sci Res Technol 2:30\u201334","journal-title":"J Sci Res Technol"},{"key":"12186_CR22","first-page":"251","volume-title":"Procedia Computer Science","author":"H Aouani","year":"2020","unstructured":"Aouani H, Ayed YB (2020) Speech Emotion Recognition with deep learning. In: Procedia Computer Science, vol 176, pp 251\u2013260"},{"key":"12186_CR23","doi-asserted-by":"crossref","unstructured":"Asiya UA, Kiran VK (2021) Speech Emotion Recognition-A Deep Learning Approach. In: Proceedings of the 5th International Conference on I-SMAC (IoT in Social, Mobile, Analytics and Cloud), I-SMAC 2021. Institute of Electrical and Electronics Engineers Inc., pp 867\u2013871","DOI":"10.1109\/I-SMAC52330.2021.9640995"},{"key":"12186_CR24","doi-asserted-by":"publisher","first-page":"1","DOI":"10.1016\/j.neucom.2023.01.002","volume":"528","author":"J de Lope","year":"2023","unstructured":"de Lope J, Gra\u00f1a M (2023) An ongoing review of speech emotion recognition. Neurocomputing 528:1\u201311. https:\/\/doi.org\/10.1016\/j.neucom.2023.01.002","journal-title":"Neurocomputing"},{"key":"12186_CR25","doi-asserted-by":"crossref","unstructured":"Elsayed N, Elsayed Z, Asadizanjani N et al (2022) Speech Emotion Recognition using Supervised Deep Recurrent System for Mental Health Monitoring. In: 2022 IEEE 8th World Forum on Internet of Things (WF-IoT), pp 1\u20136","DOI":"10.1109\/WF-IoT54382.2022.10152117"},{"key":"12186_CR26","doi-asserted-by":"publisher","unstructured":"Hashem A, Arif M, Alghamdi M (2023) Speech emotion recognition approaches: A systematic review. https:\/\/doi.org\/10.1016\/j.specom.2023.102974. Speech Communication 154:","DOI":"10.1016\/j.specom.2023.102974"},{"key":"12186_CR27","doi-asserted-by":"publisher","first-page":"1675","DOI":"10.1109\/TASLP.2021.3076364","volume":"29","author":"JH Hsu","year":"2021","unstructured":"Hsu JH, Su MH, Wu CH, Chen YH (2021) Speech Emotion Recognition Considering Nonverbal Vocalization in Affective Conversations. IEEE\/ACM Trans Audio Speech Lang Process 29:1675\u20131686. https:\/\/doi.org\/10.1109\/TASLP.2021.3076364","journal-title":"IEEE\/ACM Trans Audio Speech Lang Process"},{"key":"12186_CR28","doi-asserted-by":"publisher","unstructured":"Issa D, Fatih Demirci M, Yazici A (2020) Speech emotion recognition with deep convolutional neural networks. Biomed Signal Process Control 59. https:\/\/doi.org\/10.1016\/j.bspc.2020.101894","DOI":"10.1016\/j.bspc.2020.101894"},{"key":"12186_CR29","doi-asserted-by":"publisher","first-page":"23745","DOI":"10.1007\/s11042-020-09874-7","volume":"80","author":"R Jahangir","year":"2021","unstructured":"Jahangir R, Teh YW, Hanif F, Mujtaba G (2021) Deep learning approaches for speech emotion recognition: state of the art and research challenges. Multimedia Tools Appl 80:23745\u201323812. https:\/\/doi.org\/10.1007\/s11042-020-09874-7","journal-title":"Multimedia Tools Appl"},{"key":"12186_CR30","doi-asserted-by":"crossref","unstructured":"Li H, Zhang X, Wang MJ (2021) Research on Speech Emotion Recognition Based on Deep Neural Network. In: 2021 6th International Conference on Signal and Image Processing, ICSIP 2021. Institute of Electrical and Electronics Engineers Inc., pp 795\u2013799","DOI":"10.1109\/ICSIP52628.2021.9689043"},{"key":"12186_CR31","doi-asserted-by":"crossref","unstructured":"Pavithra A, Ledalla S, Devi JS et al (2023) Deep Learning-based Speech Emotion Recognition: An Investigation into a sustainably Emotion-Speech Relationship. In: E3S Web of Conferences. EDP Sciences","DOI":"10.1051\/e3sconf\/202343001091"},{"key":"12186_CR32","doi-asserted-by":"publisher","unstructured":"Singh J, Saheer LB, Faust O (2023) Speech Emotion Recognition Using Attention Model. Int J Environ Res Public Health 20. https:\/\/doi.org\/10.3390\/ijerph20065140","DOI":"10.3390\/ijerph20065140"},{"key":"12186_CR33","doi-asserted-by":"publisher","unstructured":"Tanko D, Dogan S, Burak Demir F et al (2022) Shoelace pattern-based speech emotion recognition of the lecturers in distance education: ShoePat23. https:\/\/doi.org\/10.1016\/j.apacoust.2022.108637. Applied Acoustics 190:","DOI":"10.1016\/j.apacoust.2022.108637"},{"key":"12186_CR34","doi-asserted-by":"publisher","unstructured":"Tyagi S, Sz\u00e9n\u00e1si S (2024) Optimizing Speech Emotion Recognition with Deep Learning and Grey Wolf Optimization: A Multi-Dataset Approach. Algorithms 17. https:\/\/doi.org\/10.3390\/a17030090","DOI":"10.3390\/a17030090"},{"key":"12186_CR35","doi-asserted-by":"publisher","unstructured":"Van LT, Le Dao TT, Le Xuan T, Castelli E (2022) Emotional Speech Recognition Using Deep Neural Networks. Sensors 22. https:\/\/doi.org\/10.3390\/s22041414","DOI":"10.3390\/s22041414"},{"key":"12186_CR36","doi-asserted-by":"publisher","unstructured":"Wang M, Ma H, Wang Y, Sun X (2024) Design of smart home system speech emotion recognition model based on ensemble deep learning and feature fusion. Appl Acoust 218. https:\/\/doi.org\/10.1016\/j.apacoust.2024.109886","DOI":"10.1016\/j.apacoust.2024.109886"},{"key":"12186_CR37","doi-asserted-by":"publisher","DOI":"10.1007\/s41060-024-00553-6","author":"A Sruthi","year":"2024","unstructured":"Sruthi A, Kumar AK, Dasari K et al (2024) Multi-language: ensemble learning-based speech emotion recognition. Int J Data Sci Analytics. https:\/\/doi.org\/10.1007\/s41060-024-00553-6","journal-title":"Int J Data Sci Analytics"},{"key":"12186_CR38","doi-asserted-by":"crossref","unstructured":"Almarzooqi NKR (2022) Speech emotion recognition framework and applications in emergency call centers. In: The 3rd International Conference on Distributed Sensing and Intelligent Systems (ICDSIS 2022). Institution of Engineering and Technology, pp 321\u2013329","DOI":"10.1049\/icp.2022.2482"},{"key":"12186_CR39","doi-asserted-by":"publisher","unstructured":"P\u0142aza M, Kaza\u0142a R, Koruba Z et al (2022) Emotion Recognition Method for Call\/Contact Centre Systems. Appl Sci (Switzerland) 12. https:\/\/doi.org\/10.3390\/app122110951","DOI":"10.3390\/app122110951"},{"key":"12186_CR40","doi-asserted-by":"crossref","unstructured":"Ahmed T, Gopala Krishnan C (2024) A Comprehensive Study on Emotion Recognition in Healthcare by Applying Machine Learning and Deep Learning Techniques. In: 2024 International Conference on Knowledge Engineering and Communication Systems (ICKECS). IEEE, Chikkaballapur, India, pp 1\u20136","DOI":"10.1109\/ICKECS61492.2024.10617024"},{"key":"12186_CR41","doi-asserted-by":"crossref","unstructured":"France DJ, Shiavi RG, Silverman S et al (2000) Acoustical Properties of Speech as Indicators of Depression and Suicidal Risk","DOI":"10.1109\/10.846676"},{"key":"12186_CR42","doi-asserted-by":"publisher","first-page":"76","DOI":"10.1016\/j.vrih.2020.10.004","volume":"3","author":"M Song","year":"2021","unstructured":"Song M, Mallol-Ragolta A, Parada-Cabaleiro E et al (2021) Frustration recognition from speech during game interaction using wide residual networks. Virtual Real Intell Hardw 3:76\u201386. https:\/\/doi.org\/10.1016\/j.vrih.2020.10.004","journal-title":"Virtual Real Intell Hardw"},{"key":"12186_CR43","doi-asserted-by":"publisher","first-page":"1888","DOI":"10.3390\/s21051888","volume":"21","author":"J Kacur","year":"2021","unstructured":"Kacur J, Puterka B, Pavlovicova J, Oravec M (2021) On the Speech Properties and Feature Extraction Methods in Speech Emotion Recognition. Sensors 21:1888. https:\/\/doi.org\/10.3390\/s21051888","journal-title":"Sensors"},{"key":"12186_CR44","doi-asserted-by":"crossref","unstructured":"Raghib O, Sharma E, Ahmad T, Alam F (2017) Emotion analysis and speech signal processing. In: 2017 IEEE International Conference on Power, Control, Signals and Instrumentation Engineering (ICPCSI). IEEE, Chennai, pp 2872\u20132875","DOI":"10.1109\/ICPCSI.2017.8392246"},{"key":"12186_CR45","doi-asserted-by":"publisher","first-page":"45","DOI":"10.1007\/s10772-020-09672-4","volume":"23","author":"A Koduru","year":"2020","unstructured":"Koduru A, Valiveti HB, Budati AK (2020) Feature extraction algorithms to improve the speech emotion recognition rate. Int J Speech Technol 23:45\u201355. https:\/\/doi.org\/10.1007\/s10772-020-09672-4","journal-title":"Int J Speech Technol"},{"key":"12186_CR46","doi-asserted-by":"publisher","first-page":"100424","DOI":"10.1016\/j.imu.2020.100424","volume":"20","author":"S Langari","year":"2020","unstructured":"Langari S, Marvi H, Zahedi M (2020) Efficient speech emotion recognition using modified feature extraction. Inf Med Unlocked 20:100424. https:\/\/doi.org\/10.1016\/j.imu.2020.100424","journal-title":"Inf Med Unlocked"},{"key":"12186_CR47","doi-asserted-by":"publisher","first-page":"1333","DOI":"10.1007\/s11277-024-11134-y","volume":"137","author":"SR Bandela","year":"2024","unstructured":"Bandela SR (2024) Teager Energy-Autocorrelation Envelope for Stressed Speech Emotion Recognition with Spectral Features: A Multi-database Analysis. Wirel Pers Commun 137:1333\u20131353. https:\/\/doi.org\/10.1007\/s11277-024-11134-y","journal-title":"Wirel Pers Commun"},{"key":"12186_CR48","doi-asserted-by":"crossref","unstructured":"Ritika CS (2024) A Comparative Review of Spectral Features in Speech- Based Emotion Recognition Techniques. In: 2024 3rd Edition of IEEE Delhi Section Flagship Conference (DELCON). IEEE, New Delhi, India, pp 1\u20137","DOI":"10.1109\/DELCON64804.2024.10866217"},{"key":"12186_CR49","doi-asserted-by":"publisher","first-page":"112880","DOI":"10.1109\/ACCESS.2025.3584534","volume":"13","author":"J Sta\u0161","year":"2025","unstructured":"Sta\u0161 J, Ond\u00e1\u0161 S, Juh\u00e1r J (2025) Performance Evaluation of Different Speech-Based Emotional Stress Level Detection Approaches. IEEE Access 13:112880\u2013112904. https:\/\/doi.org\/10.1109\/ACCESS.2025.3584534","journal-title":"IEEE Access"},{"key":"12186_CR50","doi-asserted-by":"publisher","first-page":"338","DOI":"10.1016\/j.dsp.2018.03.010","volume":"78","author":"A-O Boudraa","year":"2018","unstructured":"Boudraa A-O, Salzenstein F (2018) Teager\u2013Kaiser energy methods for signal and image analysis: A review. Digit Signal Proc 78:338\u2013375. https:\/\/doi.org\/10.1016\/j.dsp.2018.03.010","journal-title":"Digit Signal Proc"},{"key":"12186_CR51","doi-asserted-by":"publisher","first-page":"5765760","DOI":"10.1155\/2023\/5765760","volume":"2023","author":"SR Bandela","year":"2023","unstructured":"Bandela SR, Siva Priyanka S, Sunil Kumar K et al (2023) Stressed Speech Emotion Recognition Using Teager Energy and Spectral Feature Fusion with Feature Optimization. Comput Intell Neurosci 2023:5765760. https:\/\/doi.org\/10.1155\/2023\/5765760","journal-title":"Comput Intell Neurosci"},{"key":"12186_CR52","doi-asserted-by":"publisher","first-page":"1112","DOI":"10.1016\/j.protcy.2013.12.124","volume":"9","author":"JP Teixeira","year":"2013","unstructured":"Teixeira JP, Oliveira C, Lopes C (2013) Vocal Acoustic Analysis \u2013 Jitter, Shimmer and HNR Parameters. Procedia Technol 9:1112\u20131122. https:\/\/doi.org\/10.1016\/j.protcy.2013.12.124","journal-title":"Procedia Technol"},{"key":"12186_CR53","doi-asserted-by":"publisher","first-page":"6212","DOI":"10.3390\/s23136212","volume":"23","author":"R Ullah","year":"2023","unstructured":"Ullah R, Asif M, Shah WA et al (2023) Speech Emotion Recognition Using Convolution Neural Networks and Multi-Head Convolutional Transformer. Sensors 23:6212. https:\/\/doi.org\/10.3390\/s23136212","journal-title":"Sensors"},{"key":"12186_CR54","doi-asserted-by":"publisher","first-page":"13126","DOI":"10.1038\/s41598-024-63776-4","volume":"14","author":"S Akinpelu","year":"2024","unstructured":"Akinpelu S, Viriri S, Adegun A (2024) An enhanced speech emotion recognition using vision transformer. Sci Rep 14:13126. https:\/\/doi.org\/10.1038\/s41598-024-63776-4","journal-title":"Sci Rep"},{"key":"12186_CR55","doi-asserted-by":"crossref","unstructured":"Yue P, Qu L, Zheng S, Li T (2022) Multi-task Learning for Speech Emotion and Emotion Intensity Recognition. In: 2022 Asia-Pacific Signal and Information Processing Association Annual Summit and Conference (APSIPA ASC). IEEE, Chiang Mai, Thailand, pp 1232\u20131237","DOI":"10.23919\/APSIPAASC55919.2022.9979844"},{"key":"12186_CR56","doi-asserted-by":"publisher","first-page":"42","DOI":"10.1016\/j.neucom.2020.01.048","volume":"391","author":"M Hao","year":"2020","unstructured":"Hao M, Cao WH, Liu ZT et al (2020) Visual-audio emotion recognition based on multi-task and ensemble learning with multiple features. Neurocomputing 391:42\u201351. https:\/\/doi.org\/10.1016\/j.neucom.2020.01.048","journal-title":"Neurocomputing"},{"key":"12186_CR57","doi-asserted-by":"crossref","unstructured":"Wang X, Wang M, Qi W et al (2021) A Novel end-to-end Speech Emotion Recognition Network with Stacked Transformer Layers. In: ICASSP 2021\u20132021 IEEE International Conference on Acoustics, Speech and Signal Processing (ICASSP). IEEE, Toronto, ON, Canada, pp 6289\u20136293","DOI":"10.1109\/ICASSP39728.2021.9414314"},{"key":"12186_CR58","doi-asserted-by":"publisher","unstructured":"Rayhan Ahmed M, Islam S, Muzahidul Islam AKM, Shatabda S (2023) An ensemble 1D-CNN-LSTM-GRU model with data augmentation for speech emotion recognition. Expert Syst Appl 218. https:\/\/doi.org\/10.1016\/j.eswa.2023.119633","DOI":"10.1016\/j.eswa.2023.119633"},{"key":"12186_CR59","doi-asserted-by":"publisher","DOI":"10.1007\/s11042-024-19674-y","author":"KB Bhangale","year":"2024","unstructured":"Bhangale KB, Kothandaraman M (2024) A novel two-way feature extraction technique using multiple acoustic and wavelets packets for deep learning based speech emotion recognition. Multimedia Tools Appl. https:\/\/doi.org\/10.1007\/s11042-024-19674-y","journal-title":"Multimedia Tools Appl"},{"key":"12186_CR60","doi-asserted-by":"publisher","unstructured":"Geetha AV, Mala T, Priyanka D, Uma E (2024) Multimodal Emotion Recognition with Deep Learning: Advancements, challenges, and future directions. Inform Fusion 105. https:\/\/doi.org\/10.1016\/j.inffus.2023.102218","DOI":"10.1016\/j.inffus.2023.102218"},{"key":"12186_CR61","doi-asserted-by":"crossref","unstructured":"Kegkeroglou N, Filntisis PP, Maragos P (2023) Medical Face Masks and Emotion Recognition from the Body: Insights from a Deep Learning Perspective. In: ACM International Conference Proceeding Series. Association for Computing Machinery, pp 69\u201376","DOI":"10.1145\/3594806.3594829"},{"key":"12186_CR62","doi-asserted-by":"publisher","unstructured":"Lian H, Lu C, Li S et al (2023) A Survey of Deep Learning-Based Multimodal Emotion Recognition: Speech, Text, and Face. Entropy 25. https:\/\/doi.org\/10.3390\/e25101440","DOI":"10.3390\/e25101440"},{"key":"12186_CR63","doi-asserted-by":"publisher","unstructured":"Middya AI, Nag B, Roy S (2022) Deep learning based multimodal emotion recognition using model-level fusion of audio\u2013visual modalities. Knowl Based Syst 244. https:\/\/doi.org\/10.1016\/j.knosys.2022.108580","DOI":"10.1016\/j.knosys.2022.108580"},{"key":"12186_CR64","unstructured":"Rasheed BH, Yuvaraj D, Alnuaimi SS, Shanmuga Priya S (2024) International Journal of INTELLIGENT SYSTEMS AND APPLICATIONS IN ENGINEERING Automatic Speech Emotion Recognition. Using Hybrid Deep Learning Techniques"},{"key":"12186_CR65","doi-asserted-by":"publisher","unstructured":"Sarker IH (2021) Deep Learning: A Comprehensive Overview on Techniques, Taxonomy, Applications and Research Directions. SN Comput Sci 2. https:\/\/doi.org\/10.1007\/s42979-021-00815-1","DOI":"10.1007\/s42979-021-00815-1"},{"key":"12186_CR66","doi-asserted-by":"publisher","DOI":"10.1016\/j.eswa.2023.121692","author":"S Zhang","year":"2024","unstructured":"Zhang S, Yang Y, Chen C et al (2024) Deep learning-based multimodal emotion recognition from audio, visual, and text modalities: A systematic review of recent advancements and future prospects. Expert Syst Appl. https:\/\/doi.org\/10.1016\/j.eswa.2023.121692. 237:","journal-title":"Expert Syst Appl"},{"key":"12186_CR67","doi-asserted-by":"crossref","unstructured":"Krishna DN, Patil A (2020) Multimodal emotion recognition using cross-modal attention and 1D convolutional neural networks. In: Proceedings of the Annual Conference of the International Speech Communication Association, INTERSPEECH. International Speech Communication Association, pp 4243\u20134247","DOI":"10.21437\/Interspeech.2020-1190"},{"key":"12186_CR68","doi-asserted-by":"publisher","unstructured":"Luna-Jim\u00e9nez C, Griol D, Callejas Z et al (2021) Multimodal emotion recognition on RAVDESS dataset using transfer learning. Sensors 21. https:\/\/doi.org\/10.3390\/s21227665","DOI":"10.3390\/s21227665"},{"key":"12186_CR69","doi-asserted-by":"publisher","unstructured":"Page MJ, McKenzie JE, Bossuyt PM et al (2021) The PRISMA 2020 statement: an updated guideline for reporting systematic reviews. https:\/\/doi.org\/10.1136\/bmj.n71. BMJ n71","DOI":"10.1136\/bmj.n71"},{"key":"12186_CR70","doi-asserted-by":"publisher","first-page":"117327","DOI":"10.1109\/ACCESS.2019.2936124","volume":"7","author":"RA Khalil","year":"2019","unstructured":"Khalil RA, Jones E, Babar MI et al (2019) Speech Emotion Recognition Using Deep Learning Techniques: A Review. IEEE Access 7:117327\u2013117345. https:\/\/doi.org\/10.1109\/ACCESS.2019.2936124","journal-title":"IEEE Access"},{"key":"12186_CR71","doi-asserted-by":"publisher","first-page":"515","DOI":"10.1007\/s11277-023-10296-5","volume":"130","author":"AA Anthony","year":"2023","unstructured":"Anthony AA, Patil CM (2023) Speech Emotion Recognition Systems: A Comprehensive Review on Different Methodologies. Wireless Pers Commun 130:515\u2013525. https:\/\/doi.org\/10.1007\/s11277-023-10296-5","journal-title":"Wireless Pers Commun"},{"key":"12186_CR72","doi-asserted-by":"publisher","unstructured":"Pan B, Hirota K, Jia Z, Dai Y (2023) A review of multimodal emotion recognition from datasets, preprocessing, features, and fusion methods. Neurocomputing 561. https:\/\/doi.org\/10.1016\/j.neucom.2023.126866","DOI":"10.1016\/j.neucom.2023.126866"},{"key":"12186_CR73","doi-asserted-by":"publisher","first-page":"245","DOI":"10.1016\/j.neucom.2022.04.028","volume":"492","author":"YB Singh","year":"2022","unstructured":"Singh YB, Goel S (2022) A systematic literature review of speech emotion recognition approaches. Neurocomputing 492:245\u2013263. https:\/\/doi.org\/10.1016\/j.neucom.2022.04.028","journal-title":"Neurocomputing"},{"key":"12186_CR74","doi-asserted-by":"publisher","first-page":"43","DOI":"10.5120\/14977-3196","volume":"86","author":"M C.Jain","year":"2014","unstructured":"C.Jain M, Kulkarni Y V (2014) TexEmo: Conveying Emotion from Text- The Study. Int J Comput Appl 86:43\u201349. https:\/\/doi.org\/10.5120\/14977-3196","journal-title":"Int J Comput Appl"},{"key":"12186_CR75","doi-asserted-by":"publisher","first-page":"66","DOI":"10.1177\/1754073912451351","volume":"5","author":"KA Lindquist","year":"2013","unstructured":"Lindquist KA, Gendron M (2013) What\u2019s in a word? Language constructs emotion perception. Emot Rev 5:66\u201371. https:\/\/doi.org\/10.1177\/1754073912451351","journal-title":"Emot Rev"},{"key":"12186_CR76","doi-asserted-by":"publisher","first-page":"1","DOI":"10.1093\/migration\/mnad018","volume":"12","author":"J Dennison","year":"2024","unstructured":"Dennison J (2024) Emotions: functions and significance for attitudes, behaviour, and communication. Migration Stud 12:1\u201320. https:\/\/doi.org\/10.1093\/migration\/mnad018","journal-title":"Migration Stud"},{"key":"12186_CR77","doi-asserted-by":"crossref","unstructured":"Cordaro D (2024) Basic Emotion Theory: A Beginner\u2019s Guide. In: Al-Shawaf L, Shackelford TK (eds) The Oxford Handbook of Evolution and the Emotions. Oxford University Press, Oxford, pp 3\u201320","DOI":"10.1093\/oxfordhb\/9780197544754.013.1"},{"key":"12186_CR78","doi-asserted-by":"crossref","unstructured":"Geethu V, Vrindha MK, Anurenjan PR et al (2023) Speech Emotion Recognition, Datasets, Features and Models: A Review. In: 2023 International Conference on Control, Communication and Computing (ICCC). IEEE, Thiruvananthapuram, India, pp 1\u20136","DOI":"10.1109\/ICCC57789.2023.10165567"},{"key":"12186_CR79","doi-asserted-by":"crossref","unstructured":"Shareefunnisa S, Kausik CV, Kaumudi CL, Sekhar ER Delineating Emotions in Speech: Comparative Insights from Machine Learning and Deep Learning. In: 2024 15th International Conference on Computing Communication and, Technologies N (2024) (ICCCNT). IEEE, Kamand, India, pp 1\u20139","DOI":"10.1109\/ICCCNT61001.2024.10724483"},{"key":"12186_CR80","doi-asserted-by":"publisher","first-page":"2533","DOI":"10.1016\/j.procs.2023.01.227","volume":"218","author":"V Singh","year":"2023","unstructured":"Singh V, Prasad S (2023) Speech emotion recognition system using gender dependent convolution neural network. Procedia Comput Sci 218:2533\u20132540. https:\/\/doi.org\/10.1016\/j.procs.2023.01.227","journal-title":"Procedia Comput Sci"},{"key":"12186_CR81","doi-asserted-by":"publisher","first-page":"310","DOI":"10.4324\/9781315559940-18","volume-title":"Emotion Theory: The Routledge Comprehensive Guide","author":"MN Shiota","year":"2024","unstructured":"Shiota MN (2024) Basic and Discrete Emotion Theories. In: Scarantino A (eds) Emotion Theory: The Routledge Comprehensive Guide, Volume I: History, Contemporary Theories, and Key Elements. Routledge, New York, pp 310\u2013330"},{"key":"12186_CR82","doi-asserted-by":"publisher","unstructured":"Dong B, Xu G (2022) An Empirical Study on the Evaluation of Emotional Complexity in Daily Life. Front Psychol 13. https:\/\/doi.org\/10.3389\/fpsyg.2022.839133","DOI":"10.3389\/fpsyg.2022.839133"},{"key":"12186_CR83","doi-asserted-by":"publisher","first-page":"215","DOI":"10.3389\/fpsyg.2018.00215","volume":"9","author":"J Cespedes-Guevara","year":"2018","unstructured":"Cespedes-Guevara J, Eerola T (2018) Music Communicates Affects, Not Basic Emotions \u2013 A Constructionist Account of Attribution of Emotional Meanings to Music. Front Psychol 9:215. https:\/\/doi.org\/10.3389\/fpsyg.2018.00215","journal-title":"Front Psychol"},{"key":"12186_CR84","doi-asserted-by":"crossref","unstructured":"Wu H, Chou H-C, Chang K-W et al (2024) Empower Typed Descriptions by Large Language Models for Speech Emotion Recognition. In: 2024 Asia Pacific Signal and Information Processing Association Annual Summit and Conference (APSIPA ASC). IEEE, Macau, Macao, pp 1\u20136","DOI":"10.1109\/APSIPAASC63619.2025.10848758"},{"key":"12186_CR85","doi-asserted-by":"crossref","unstructured":"Lea\u00f1o JM, Stewart CAC, Pe\u00f1a CF (2023) Multilabel Emotion Recognition Through Sequence Labeling and Sentence Classification Models using Textual Data. In: 2023 IEEE International Conference on Communication, Networks and Satellite (COMNETSAT). IEEE, Malang, Indonesia, pp 378\u2013385","DOI":"10.1109\/COMNETSAT59769.2023.10420660"},{"key":"12186_CR86","doi-asserted-by":"publisher","unstructured":"Vago DR, Silbersweig DA (2012) Self-awareness, self-regulation, and self-transcendence (S-ART): a framework for understanding the neurobiological mechanisms of mindfulness. Front Hum Neurosci 6. https:\/\/doi.org\/10.3389\/fnhum.2012.00296","DOI":"10.3389\/fnhum.2012.00296"},{"key":"12186_CR87","doi-asserted-by":"publisher","DOI":"10.4324\/9781003298076","volume-title":"The Complexity of Trauma: Jungian and Psychoanalytic Approaches to the Treatment of Trauma","author":"L Zoppi","year":"2024","unstructured":"Zoppi L, Schmidt M (2024) The Complexity of Trauma: Jungian and Psychoanalytic Approaches to the Treatment of Trauma, 1st edn. Routledge, London","edition":"1"},{"key":"12186_CR88","doi-asserted-by":"publisher","first-page":"117","DOI":"10.3390\/bs13020117","volume":"13","author":"M Majolo","year":"2023","unstructured":"Majolo M, Gomes WB, DeCastro TG (2023) Self-Consciousness and Self-Awareness: Associations between Stable and Transitory Levels of Evidence. Behav Sci 13:117. https:\/\/doi.org\/10.3390\/bs13020117","journal-title":"Behav Sci"},{"key":"12186_CR89","doi-asserted-by":"publisher","first-page":"281","DOI":"10.1080\/02699938808412701","volume":"2","author":"RS Lazarus","year":"1988","unstructured":"Lazarus RS, Smith CA (1988) Knowledge and Appraisal in the Cognition\u2014Emotion Relationship. Cognition Emot 2:281\u2013300. https:\/\/doi.org\/10.1080\/02699938808412701","journal-title":"Cognition Emot"},{"key":"12186_CR90","doi-asserted-by":"publisher","first-page":"54","DOI":"10.1177\/0963721411430832","volume":"21","author":"WA Cunningham","year":"2012","unstructured":"Cunningham WA, Brosch T (2012) Motivational Salience: Amygdala Tuning From Traits, Needs, Values, and Goals. Curr Dir Psychol Sci 21:54\u201359. https:\/\/doi.org\/10.1177\/0963721411430832","journal-title":"Curr Dir Psychol Sci"},{"key":"12186_CR91","doi-asserted-by":"publisher","first-page":"5506","DOI":"10.3390\/s24175506","volume":"24","author":"H Li","year":"2024","unstructured":"Li H, Li J, Liu H et al (2024) MelTrans: Mel-Spectrogram Relationship-Learning for Speech Emotion Recognition via Transformers. Sensors 24:5506. https:\/\/doi.org\/10.3390\/s24175506","journal-title":"Sensors"},{"key":"12186_CR92","doi-asserted-by":"crossref","unstructured":"Plutchik R (1961) Studies of Emotion in the Light of a New Theory","DOI":"10.2466\/pr0.1961.8.1.170"},{"key":"12186_CR93","doi-asserted-by":"publisher","first-page":"408","DOI":"10.1109\/TAFFC.2019.2945322","volume":"13","author":"Z Xiao","year":"2022","unstructured":"Xiao Z, Chen Y, Dou W et al (2022) MES-P: An Emotional Tonal Speech Dataset in Mandarin with Distal and Proximal Labels. IEEE Trans Affect Comput 13:408\u2013425. https:\/\/doi.org\/10.1109\/TAFFC.2019.2945322","journal-title":"IEEE Trans Affect Comput"},{"key":"12186_CR94","doi-asserted-by":"crossref","unstructured":"Cha J-H, Kim S-B, Oh H-S, Lee S-W (2025) JELLY: Joint Emotion Recognition and Context Reasoning with LLMs for Conversational Speech Synthesis. In: ICASSP 2025\u20132025 IEEE International Conference on Acoustics, Speech and Signal Processing (ICASSP). IEEE, Hyderabad, India, pp 1\u20135","DOI":"10.1109\/ICASSP49660.2025.10888227"},{"key":"12186_CR95","doi-asserted-by":"crossref","unstructured":"Wu Y, Yue P, Cheng C, Li T (2023) Investigation of Ensemble of Self-Supervised Models for Speech Emotion Recognition. In: 2023 Asia Pacific Signal and Information Processing Association Annual Summit and Conference (APSIPA ASC). IEEE, Taipei, Taiwan, pp 988\u2013995","DOI":"10.1109\/APSIPAASC58517.2023.10317508"},{"key":"12186_CR96","doi-asserted-by":"publisher","first-page":"3177","DOI":"10.1109\/TAFFC.2023.3253859","volume":"14","author":"YC Wu","year":"2023","unstructured":"Wu YC, Chiu LW, Lai CC et al (2023) Recognizing, Fast and Slow: Complex Emotion Recognition With Facial Expression Detection and Remote Physiological Measurement. IEEE Trans Affect Comput 14:3177\u20133190. https:\/\/doi.org\/10.1109\/TAFFC.2023.3253859","journal-title":"IEEE Trans Affect Comput"},{"key":"12186_CR97","doi-asserted-by":"publisher","first-page":"781","DOI":"10.3389\/fpsyg.2019.00781","volume":"10","author":"S Gu","year":"2019","unstructured":"Gu S, Wang F, Patel NP et al (2019) A Model for Basic Emotions Using Observations of Behavior in Drosophila. Front Psychol 10:781. https:\/\/doi.org\/10.3389\/fpsyg.2019.00781","journal-title":"Front Psychol"},{"key":"12186_CR98","doi-asserted-by":"crossref","unstructured":"Burkhardt F, Paeschke A, Rolfes M et al (2005) A database of German emotional speech. In: 9th European Conference on Speech Communication and Technology. pp 1517\u20131520","DOI":"10.21437\/Interspeech.2005-446"},{"key":"12186_CR99","doi-asserted-by":"publisher","DOI":"10.5281\/zenodo.1188976","author":"SR Livingstone","year":"2018","unstructured":"Livingstone SR, Russo FA (2018) The Ryerson Audio-Visual Database of Emotional Speech and Song (RAVDESS): A dynamic, multimodal set of facial and vocal expressions in North American English. PLoS ONE. https:\/\/doi.org\/10.5281\/zenodo.1188976","journal-title":"PLoS ONE"},{"key":"12186_CR100","doi-asserted-by":"publisher","unstructured":"Perepelkina O, Konstantinova M, Kazimirova E (2018) RAMAS: Russian Multimodal Corpus of Dyadic Interaction for. https:\/\/doi.org\/10.7287\/peerj.preprints.26688v1. Studying Emotion Recognition","DOI":"10.7287\/peerj.preprints.26688v1"},{"key":"12186_CR101","doi-asserted-by":"publisher","unstructured":"Pichora-Fuller P MK, Dupuis K (2020) Toronto emotional speech set (TESS). https:\/\/doi.org\/10.5683\/SP2\/E8H2MF","DOI":"10.5683\/SP2\/E8H2MF"},{"key":"12186_CR102","doi-asserted-by":"publisher","first-page":"335","DOI":"10.1007\/s10579-008-9076-6","volume":"42","author":"C Busso","year":"2008","unstructured":"Busso C, Bulut M, Lee CC et al (2008) IEMOCAP: Interactive emotional dyadic motion capture database. Lang Resour Evaluation 42:335\u2013359. https:\/\/doi.org\/10.1007\/s10579-008-9076-6","journal-title":"Lang Resour Evaluation"},{"key":"12186_CR103","doi-asserted-by":"publisher","first-page":"913","DOI":"10.1007\/s12652-016-0406-z","volume":"8","author":"Y Li","year":"2017","unstructured":"Li Y, Tao J, Chao L et al (2017) CHEAVD: a Chinese natural emotional audio\u2013visual database. J Ambient Intell Humaniz Comput 8:913\u2013924. https:\/\/doi.org\/10.1007\/s12652-016-0406-z","journal-title":"J Ambient Intell Humaniz Comput"},{"key":"12186_CR104","doi-asserted-by":"crossref","unstructured":"Martin O, Kotsia I, Macq B, Pitas I (2006) The eNTERFACE\u201905. Audio-Visual Emotion Database","DOI":"10.1109\/ICDEW.2006.145"},{"key":"12186_CR105","doi-asserted-by":"crossref","unstructured":"Nojavanasghari B, Baltru\u0161aitis T, Hughes CE, Morency LP (2016) Emo react: A multimodal approach and dataset for recognizing emotional responses in children. In: ICMI 2016 - Proceedings of the 18th ACM International Conference on Multimodal Interaction. Association for Computing Machinery, Inc, pp 137\u2013144","DOI":"10.1145\/2993148.2993168"},{"key":"12186_CR106","doi-asserted-by":"crossref","unstructured":"Poria S, Hazarika D, Majumder N et al (2019) MELD: A Multimodal Multi-Party Dataset for Emotion Recognition in Conversations","DOI":"10.18653\/v1\/P19-1050"},{"key":"12186_CR107","unstructured":"Zadeh A, Liang PP, Vanbriesen J et al (2018) Multimodal Language Analysis in the Wild: CMU-MOSEI Dataset and Interpretable Dynamic Fusion Graph. Association for Computational Linguistics"},{"key":"12186_CR108","doi-asserted-by":"crossref","unstructured":"Haq S, Jackson PJB (2010) Multimodal emotion recognition. In: Machine Audition: Principles, Algorithms and Systems. IGI Global, pp 398\u2013423","DOI":"10.4018\/978-1-61520-919-4.ch017"},{"key":"12186_CR109","doi-asserted-by":"publisher","first-page":"377","DOI":"10.1109\/TAFFC.2014.2336244","volume":"5","author":"H Cao","year":"2014","unstructured":"Cao H, Cooper DG, Keutmann MK et al (2014) CREMA-D: Crowd-sourced emotional multimodal actors dataset. IEEE Trans Affect Comput 5:377\u2013390. https:\/\/doi.org\/10.1109\/TAFFC.2014.2336244","journal-title":"IEEE Trans Affect Comput"},{"key":"12186_CR110","doi-asserted-by":"publisher","first-page":"28373","DOI":"10.1007\/s11042-023-16443-1","volume":"83","author":"P Kumar","year":"2024","unstructured":"Kumar P, Malik S, Raman B (2024) Interpretable multimodal emotion recognition using hybrid fusion of speech and image data. Multimedia Tools Appl 83:28373\u201328394. https:\/\/doi.org\/10.1007\/s11042-023-16443-1","journal-title":"Multimedia Tools Appl"},{"key":"12186_CR111","doi-asserted-by":"publisher","first-page":"67","DOI":"10.1109\/TAFFC.2016.2515617","volume":"8","author":"C Busso","year":"2017","unstructured":"Busso C, Parthasarathy S, Burmania A et al (2017) MSP-IMPROV: An Acted Corpus of Dyadic Interactions to Study Emotion Perception. IEEE Trans Affect Comput 8:67\u201380. https:\/\/doi.org\/10.1109\/TAFFC.2016.2515617","journal-title":"IEEE Trans Affect Comput"},{"key":"12186_CR112","doi-asserted-by":"publisher","first-page":"1","DOI":"10.1145\/3529759","volume":"22","author":"EA Retta","year":"2023","unstructured":"Retta EA, Almekhlafi E, Sutcliffe R et al (2023) A New Amharic Speech Emotion Dataset and Classification Benchmark. ACM Trans Asian Low-Resour Lang Inf Process 22:1\u201322. https:\/\/doi.org\/10.1145\/3529759","journal-title":"ACM Trans Asian Low-Resour Lang Inf Process"},{"key":"12186_CR113","doi-asserted-by":"publisher","first-page":"5264","DOI":"10.3758\/s13428-023-02270-7","volume":"56","author":"CS Chong","year":"2023","unstructured":"Chong CS, Davis C, Kim J (2023) A Cantonese Audio-Visual Emotional Speech (CAVES) dataset. Behav Res 56:5264\u20135278. https:\/\/doi.org\/10.3758\/s13428-023-02270-7","journal-title":"Behav Res"},{"key":"12186_CR114","doi-asserted-by":"publisher","first-page":"1142","DOI":"10.1109\/TASLPRO.2025.3540662","volume":"33","author":"F Catania","year":"2025","unstructured":"Catania F, Wilke JW, Garzotto F (2025) Emozionalmente: A Crowdsourced Corpus of Simulated Emotional Speech in Italian. IEEE Trans Audio Speech Lang Process 33:1142\u20131155. https:\/\/doi.org\/10.1109\/TASLPRO.2025.3540662","journal-title":"IEEE Trans Audio Speech Lang Process"},{"key":"12186_CR115","doi-asserted-by":"publisher","first-page":"4509","DOI":"10.3390\/app15084509","volume":"15","author":"Y Fu","year":"2025","unstructured":"Fu Y, Liu Q, Song Q et al (2025) Multi-HM: A Chinese Multimodal Dataset and Fusion Framework for Emotion Recognition in Human\u2013Machine Dialogue Systems. Appl Sci 15:4509. https:\/\/doi.org\/10.3390\/app15084509","journal-title":"Appl Sci"},{"key":"12186_CR116","doi-asserted-by":"publisher","first-page":"13093","DOI":"10.1007\/s11042-023-15959-w","volume":"83","author":"E Garcia-Cuesta","year":"2023","unstructured":"Garcia-Cuesta E, Salvador AB, P\u00e3ez DG (2023) EmoMatchSpanishDB: study of speech emotion recognition machine learning models in a new Spanish elicited database. Multimed Tools Appl 83:13093\u201313112. https:\/\/doi.org\/10.1007\/s11042-023-15959-w","journal-title":"Multimed Tools Appl"},{"key":"12186_CR117","doi-asserted-by":"crossref","unstructured":"Fan W, Xu X, Xing X et al (2021) LSSED: A large-scale dataset and benchmark for speech emotion recognition. In: ICASSP, IEEE International Conference on Acoustics, Speech and Signal Processing - Proceedings. Institute of Electrical and Electronics Engineers Inc., pp 641\u2013645","DOI":"10.1109\/ICASSP39728.2021.9414542"},{"key":"12186_CR118","doi-asserted-by":"crossref","unstructured":"Liu Y, Dai W, Feng C et al (2022) MAFW: A Large-scale, Multi-modal, Compound Affective Database for Dynamic Facial Expression Recognition in the Wild. In: MM 2022 - Proceedings of the 30th ACM International Conference on Multimedia. Association for Computing Machinery, Inc, pp 24\u201332","DOI":"10.1145\/3503161.3548190"},{"key":"12186_CR119","doi-asserted-by":"publisher","first-page":"3029","DOI":"10.1007\/s10579-025-09845-0","volume":"59","author":"X Wu","year":"2025","unstructured":"Wu X, Song C, Xiang S et al (2025) A Chinese natural speech complex emotion dataset based on emotion vector annotation method. Lang Resour Evaluation 59:3029\u20133050. https:\/\/doi.org\/10.1007\/s10579-025-09845-0","journal-title":"Lang Resour Evaluation"},{"key":"12186_CR120","unstructured":"Latif S, Qayyum A, Usman M, Qadir J (2020) Cross Lingual Speech Emotion Recognition: Urdu vs. Western Languages"},{"key":"12186_CR121","doi-asserted-by":"publisher","first-page":"111627","DOI":"10.1016\/j.dib.2025.111627","volume":"60","author":"MZ Akhtar","year":"2025","unstructured":"Akhtar MZ, Jahangir R, Ain Q et al (2025) UrduSER: A comprehensive dataset for speech emotion recognition in Urdu language. Data Brief 60:111627. https:\/\/doi.org\/10.1016\/j.dib.2025.111627","journal-title":"Data Brief"},{"key":"12186_CR122","doi-asserted-by":"crossref","unstructured":"Nagarajan B, Ramana V, Oruganti M (2019) Cross-Domain Transfer Learning for Complex Emotion Recognition","DOI":"10.1109\/TENSYMP46218.2019.8971023"},{"key":"12186_CR123","doi-asserted-by":"publisher","first-page":"1397","DOI":"10.1007\/s41870-023-01697-7","volume":"16","author":"PS Tomar","year":"2024","unstructured":"Tomar PS, Mathur K, Suman U (2024) Fusing facial and speech cues for enhanced multimodal emotion recognition. Int J Inform Technol (Singapore) 16:1397\u20131405. https:\/\/doi.org\/10.1007\/s41870-023-01697-7","journal-title":"Int J Inform Technol (Singapore)"},{"key":"12186_CR124","doi-asserted-by":"publisher","first-page":"4203","DOI":"10.32604\/cmc.2023.028291","volume":"74","author":"P Gong","year":"2023","unstructured":"Gong P, Liu J, Wu Z, Computers et al (2023) Mater Continua 74:4203\u20134220. https:\/\/doi.org\/10.32604\/cmc.2023.028291","journal-title":"Mater Continua"},{"key":"12186_CR125","doi-asserted-by":"publisher","first-page":"61672","DOI":"10.1109\/ACCESS.2020.2984368","volume":"8","author":"NH Ho","year":"2020","unstructured":"Ho NH, Yang HJ, Kim SH, Lee G (2020) Multimodal Approach of Speech Emotion Recognition Using Multi-Level Multi-Head Fusion Attention-Based Recurrent Neural Network. IEEE Access 8:61672\u201361686. https:\/\/doi.org\/10.1109\/ACCESS.2020.2984368","journal-title":"IEEE Access"},{"key":"12186_CR126","doi-asserted-by":"publisher","first-page":"32265","DOI":"10.1007\/s11042-022-13091-9","volume":"81","author":"N Jia","year":"2022","unstructured":"Jia N, Zheng C, Sun W (2022) A multimodal emotion recognition model integrating speech, video and MoCAP. Multimedia Tools Appl 81:32265\u201332286. https:\/\/doi.org\/10.1007\/s11042-022-13091-9","journal-title":"Multimedia Tools Appl"},{"key":"12186_CR127","doi-asserted-by":"publisher","DOI":"10.1007\/s10489-024-05630-8","author":"J Lei","year":"2024","unstructured":"Lei J, Wang J, Wang Y (2024) Multi-level attention fusion network assisted by relative entropy alignment for multimodal speech emotion recognition. Appl Intell. https:\/\/doi.org\/10.1007\/s10489-024-05630-8","journal-title":"Appl Intell"},{"key":"12186_CR128","doi-asserted-by":"publisher","unstructured":"Wang Y, Gu Y, Yin Y et al (2023) Multimodal transformer augmented fusion for speech emotion recognition. Front Neurorobotics 17. https:\/\/doi.org\/10.3389\/fnbot.2023.1181598","DOI":"10.3389\/fnbot.2023.1181598"},{"key":"12186_CR129","doi-asserted-by":"crossref","unstructured":"Chen C (2020) Research of Tri-modal Mandarin Emotion Recognition Based on Speech, Facial Expression and Body Gesture. In: 2020 International Conference on Virtual Reality and Intelligent Systems (ICVRIS). IEEE, pp 72\u201377","DOI":"10.1109\/ICVRIS51417.2020.00025"},{"key":"12186_CR130","doi-asserted-by":"crossref","unstructured":"Chen J, Liu C, Li M Automatic emotional spoken language text corpus construction from written dialogs in fictions. In: 2017 Seventh International Conference on Affective Computing and, Interaction I (2017) (ACII). IEEE, pp 319\u2013324","DOI":"10.1109\/ACII.2017.8273619"},{"key":"12186_CR131","doi-asserted-by":"crossref","unstructured":"Miao H, Zhang Y, Li W et al (2018) Chinese Multimodal Emotion Recognition in Deep and Traditional Machine Leaming Approaches. In: 2018 First Asian Conference on Affective Computing and Intelligent Interaction (ACII Asia). IEEE, pp 1\u20136","DOI":"10.1109\/ACIIAsia.2018.8470379"},{"key":"12186_CR132","doi-asserted-by":"crossref","unstructured":"Ryumina E, Verkholyak O, Karpov A (2021) Annotation confidence vs. training sample size: Trade-off solution for partially-continuous categorical emotion recognition. In: Proceedings of the Annual Conference of the International Speech Communication Association, INTERSPEECH. International Speech Communication Association, pp 4361\u20134365","DOI":"10.21437\/Interspeech.2021-1636"},{"key":"12186_CR133","doi-asserted-by":"publisher","first-page":"80","DOI":"10.22667\/JISIS.2021.02.28.080","volume":"11","author":"O Verkholyak","year":"2021","unstructured":"Verkholyak O, Dvoynikova A, Karpov A (2021) A bimodal approach for speech emotion recognition using audio and text. J Internet Serv Inform Secur 11:80\u201396. https:\/\/doi.org\/10.22667\/JISIS.2021.02.28.080","journal-title":"J Internet Serv Inform Secur"},{"key":"12186_CR134","doi-asserted-by":"publisher","unstructured":"Xie B, Sidulova M, Park CH (2021) Article robust multimodal emotion recognition from conversation with transformer-based crossmodality the title fusion. Sensors 21. https:\/\/doi.org\/10.3390\/s21144913","DOI":"10.3390\/s21144913"},{"key":"12186_CR135","doi-asserted-by":"publisher","unstructured":"Chen W, Xing X, Xu X et al (2021) Key-Sparse Transformer for Multimodal. https:\/\/doi.org\/10.1109\/ICASSP43922.2022.9746598. Speech Emotion Recognition","DOI":"10.1109\/ICASSP43922.2022.9746598"},{"key":"12186_CR136","doi-asserted-by":"publisher","first-page":"1803","DOI":"10.1109\/TASLP.2022.3171965","volume":"30","author":"W Fan","year":"2022","unstructured":"Fan W, Xu X, Cai B, Xing X (2022) ISNet: Individual Standardization Network for Speech Emotion Recognition. IEEE\/ACM Trans Audio Speech Lang Process 30:1803\u20131814. https:\/\/doi.org\/10.1109\/TASLP.2022.3171965","journal-title":"IEEE\/ACM Trans Audio Speech Lang Process"},{"key":"12186_CR137","doi-asserted-by":"crossref","unstructured":"Chen H, Huang H, Dong J et al (2024) FineCLIPER: Multi-modal Fine-grained CLIP for Dynamic Facial Expression Recognition with AdaptERs. In: Proceedings of the 32nd ACM International Conference on Multimedia. ACM, New York, NY, USA, pp 2301\u20132310","DOI":"10.1145\/3664647.3680827"},{"key":"12186_CR138","doi-asserted-by":"crossref","unstructured":"Foteinopoulou NM, Patras I (2024) EmoCLIP: A Vision-Language Method for Zero-Shot Video Facial Expression Recognition. In: 2024 IEEE 18th International Conference on Automatic Face and Gesture Recognition (FG). IEEE, pp 1\u201310","DOI":"10.1109\/FG59268.2024.10581982"},{"key":"12186_CR139","doi-asserted-by":"crossref","unstructured":"Sun L, Lian Z, Liu B, Tao J (2023) MAE-DFER: Efficient Masked Autoencoder for Self-supervised Dynamic Facial Expression Recognition. In: MM 2023 - Proceedings of the 31st ACM International Conference on Multimedia. Association for Computing Machinery, Inc, pp 6110\u20136121","DOI":"10.1145\/3581783.3612365"},{"key":"12186_CR140","doi-asserted-by":"publisher","first-page":"4340","DOI":"10.3390\/app15084340","volume":"15","author":"A Mares","year":"2025","unstructured":"Mares A, Diaz-Arango G, Perez-Jacome-Friscione J et al (2025) Advancing Spanish Speech Emotion Recognition: A Comprehensive Benchmark of Pre-Trained Models. Appl Sci 15:4340. https:\/\/doi.org\/10.3390\/app15084340","journal-title":"Appl Sci"},{"key":"12186_CR141","doi-asserted-by":"crossref","unstructured":"Chen Y, Fan W, Xing X et al (2022) CPED: A Large-Scale Chinese Personalized. and Emotional Dialogue Dataset for Conversational AI","DOI":"10.36227\/techrxiv.19919483"},{"key":"12186_CR142","doi-asserted-by":"publisher","unstructured":"He H, Li B, Xiong Y et al (2025) Heuristic personality recognition based on fusing multiple conversations and utterance-level affection. Inf Process Manage 62. https:\/\/doi.org\/10.1016\/j.ipm.2024.103931","DOI":"10.1016\/j.ipm.2024.103931"},{"key":"12186_CR143","doi-asserted-by":"publisher","first-page":"12587","DOI":"10.3390\/app132312587","volume":"13","author":"EA Retta","year":"2023","unstructured":"Retta EA, Sutcliffe R, Mahmood J et al (2023) Cross-Corpus Multilingual Speech Emotion Recognition: Amharic vs. Other Languages. Appl Sci 13:12587. https:\/\/doi.org\/10.3390\/app132312587","journal-title":"Appl Sci"},{"key":"12186_CR144","doi-asserted-by":"publisher","first-page":"2563","DOI":"10.1093\/cercor\/bhv086","volume":"26","author":"H Saarim\u00e4ki","year":"2016","unstructured":"Saarim\u00e4ki H, Gotsopoulos A, J\u00e4\u00e4skel\u00e4inen IP et al (2016) Discrete Neural Signatures of Basic Emotions. Cereb Cortex 26:2563\u20132573. https:\/\/doi.org\/10.1093\/cercor\/bhv086","journal-title":"Cereb Cortex"},{"key":"12186_CR145","doi-asserted-by":"publisher","first-page":"1432","DOI":"10.3389\/fpsyg.2017.01432","volume":"8","author":"A Celeghin","year":"2017","unstructured":"Celeghin A, Diano M, Bagnis A et al (2017) Basic Emotions in Human Neuroscience: Neuroimaging and Beyond. Front Psychol 8:1432. https:\/\/doi.org\/10.3389\/fpsyg.2017.01432","journal-title":"Front Psychol"},{"key":"12186_CR146","doi-asserted-by":"publisher","first-page":"28","DOI":"10.1007\/s10462-023-10662-6","volume":"57","author":"K Berahmand","year":"2024","unstructured":"Berahmand K, Daneshfar F, Salehi ES et al (2024) Autoencoders and their applications in machine learning: a survey. Artif Intell Rev 57:28. https:\/\/doi.org\/10.1007\/s10462-023-10662-6","journal-title":"Artif Intell Rev"},{"key":"12186_CR147","doi-asserted-by":"publisher","first-page":"895","DOI":"10.1016\/j.procs.2018.04.298","volume":"131","author":"G Shen","year":"2018","unstructured":"Shen G, Tan Q, Zhang H et al (2018) Deep Learning with Gated Recurrent Unit Networks for Financial Sequence Predictions. Procedia Comput Sci 131:895\u2013903. https:\/\/doi.org\/10.1016\/j.procs.2018.04.298","journal-title":"Procedia Comput Sci"},{"key":"12186_CR148","doi-asserted-by":"crossref","unstructured":"Avs KP, Saksena A (2024) S Decoding Emotions: A Deep Learning Framework for Speech Emotion Recognition. In: 2024 International Conference on Smart Technologies for Sustainable Development Goals (ICSTSDG). IEEE, Chennai\u2009\u2013\u2009600077, Tamil Nadu, India, pp 1\u20138","DOI":"10.1109\/ICSTSDG61998.2024.11026432"},{"key":"12186_CR149","doi-asserted-by":"publisher","unstructured":"Kim S, Lee SP (2023) A BiLSTM\u2013Transformer and 2D CNN Architecture for Emotion Recognition from Speech. Electron (Switzerland) 12. https:\/\/doi.org\/10.3390\/electronics12194034","DOI":"10.3390\/electronics12194034"},{"key":"12186_CR150","doi-asserted-by":"publisher","first-page":"188","DOI":"10.1109\/TNNLS.2023.3304516","volume":"36","author":"L Guo","year":"2025","unstructured":"Guo L, Ding S, Wang L, Dang J (2025) DSTCNet: Deep Spectro-Temporal-Channel Attention Network for Speech Emotion Recognition. IEEE Trans Neural Netw Learn Syst 36:188\u2013197. https:\/\/doi.org\/10.1109\/TNNLS.2023.3304516","journal-title":"IEEE Trans Neural Netw Learn Syst"},{"key":"12186_CR151","doi-asserted-by":"publisher","unstructured":"Zheng C, Wang C, Jia N (2020) An ensemble model for multi-level speech emotion recognition. Appl Sci (Switzerland) 10. https:\/\/doi.org\/10.3390\/app10010205","DOI":"10.3390\/app10010205"},{"key":"12186_CR152","doi-asserted-by":"publisher","DOI":"10.1109\/TAFFC.2024.3369726","author":"W Chen","year":"2024","unstructured":"Chen W, Xing X, Chen P, Xu X (2024) Vesper: A Compact and Effective Pretrained Model for Speech Emotion Recognition. IEEE Trans Affect Comput. https:\/\/doi.org\/10.1109\/TAFFC.2024.3369726","journal-title":"IEEE Trans Affect Comput"},{"key":"12186_CR153","doi-asserted-by":"publisher","first-page":"121488","DOI":"10.1016\/j.ins.2024.121488","volume":"689","author":"CP Udeh","year":"2025","unstructured":"Udeh CP, Chen L, Du S et al (2025) Improved ShuffleNet V2 network with attention for speech emotion recognition. Inf Sci 689:121488. https:\/\/doi.org\/10.1016\/j.ins.2024.121488","journal-title":"Inf Sci"},{"key":"12186_CR154","doi-asserted-by":"publisher","unstructured":"Caulley D, Alemu Y, Burson S et al (2023) Objectively Quantifying Pediatric Psychiatric Severity Using Artificial Intelligence, Voice Recognition Technology, and Universal Emotions: Pilot Study for Artificial Intelligence-Enabled Innovation to Address Youth Mental Health Crisis. JMIR Res Protocols 12. https:\/\/doi.org\/10.2196\/51912","DOI":"10.2196\/51912"},{"key":"12186_CR155","doi-asserted-by":"publisher","unstructured":"Khan M, Gueaieb W, El Saddik A, Kwon S (2024) MSER: Multimodal speech emotion recognition using cross-attention with deep fusion. Expert Syst Appl 245. https:\/\/doi.org\/10.1016\/j.eswa.2023.122946","DOI":"10.1016\/j.eswa.2023.122946"},{"key":"12186_CR156","doi-asserted-by":"publisher","first-page":"5473","DOI":"10.1038\/s41598-025-89202-x","volume":"15","author":"M Khan","year":"2025","unstructured":"Khan M, Tran P-N, Pham NT et al (2025) MemoCMT: multimodal emotion recognition using cross-modal transformer-based feature fusion. Sci Rep 15:5473. https:\/\/doi.org\/10.1038\/s41598-025-89202-x","journal-title":"Sci Rep"},{"key":"12186_CR157","doi-asserted-by":"publisher","first-page":"31","DOI":"10.1007\/s11227-024-06582-z","volume":"81","author":"Z Jin","year":"2025","unstructured":"Jin Z, Zai W (2025) Audiovisual emotion recognition based on bi-layer LSTM and multi-head attention mechanism on RAVDESS dataset. J Supercomput 81:31. https:\/\/doi.org\/10.1007\/s11227-024-06582-z","journal-title":"J Supercomput"},{"key":"12186_CR158","doi-asserted-by":"publisher","first-page":"1092","DOI":"10.1109\/TCE.2025.3532322","volume":"71","author":"M Khan","year":"2025","unstructured":"Khan M, Ahmad J, Gueaieb W et al (2025) Joint Multi-Scale Multimodal Transformer for Emotion Using Consumer Devices. IEEE Trans Consumer Electron 71:1092\u20131101. https:\/\/doi.org\/10.1109\/TCE.2025.3532322","journal-title":"IEEE Trans Consumer Electron"},{"key":"12186_CR159","doi-asserted-by":"publisher","unstructured":"Haque MDR, Rubya S (2023) An Overview of Chatbot-Based Mobile Mental Health Apps: Insights From App Description and User Reviews. JMIR mHealth and uHealth 11. https:\/\/doi.org\/10.2196\/44838","DOI":"10.2196\/44838"},{"key":"12186_CR160","doi-asserted-by":"crossref","unstructured":"Rammohan RA, Medikonda J, Pothiyil DI (2020) Speech Signal-Based Modelling of Basic Emotions to Analyse Compound Emotion: Anxiety. In: 2020 IEEE International Conference on Distributed Computing, VLSI, Electrical Circuits and Robotics (DISCOVER). IEEE, Udupi, India, pp 218\u2013223","DOI":"10.1109\/DISCOVER50404.2020.9278094"},{"key":"12186_CR161","doi-asserted-by":"publisher","unstructured":"Wang H, Kim DH (2024) Graph Neural Network-Based Speech Emotion Recognition: A Fusion of Skip Graph Convolutional Networks and Graph Attention Networks. Electron (Switzerland) 13. https:\/\/doi.org\/10.3390\/electronics13214208","DOI":"10.3390\/electronics13214208"},{"key":"12186_CR162","doi-asserted-by":"publisher","first-page":"1","DOI":"10.3390\/s21041249","volume":"21","author":"BJ Abbaschian","year":"2021","unstructured":"Abbaschian BJ, Sierra-Sosa D, Elmaghraby A (2021) Deep learning techniques for speech emotion recognition, from databases to models. Sens (Switzerland) 21:1\u201327. https:\/\/doi.org\/10.3390\/s21041249","journal-title":"Sens (Switzerland)"},{"key":"12186_CR163","doi-asserted-by":"publisher","first-page":"969","DOI":"10.1109\/TAFFC.2024.3485057","volume":"16","author":"W-B Jiang","year":"2025","unstructured":"Jiang W-B, Liu X-H, Zheng W-L, Lu B-L (2025) SEED-VII: A Multimodal Dataset of Six Basic Emotions With Continuous Labels for Emotion Recognition. IEEE Trans Affect Comput 16:969\u2013985. https:\/\/doi.org\/10.1109\/TAFFC.2024.3485057","journal-title":"IEEE Trans Affect Comput"},{"key":"12186_CR164","doi-asserted-by":"crossref","unstructured":"Yue P, Zheng S, Li T (2023) Complex Feature Information Enhanced Speech Emotion Recognition. In: 2023 Asia Pacific Signal and Information Processing Association Annual Summit and Conference, APSIPA ASC 2023. Institute of Electrical and Electronics Engineers Inc., pp 941\u2013946","DOI":"10.1109\/APSIPAASC58517.2023.10317348"},{"key":"12186_CR165","doi-asserted-by":"publisher","first-page":"966","DOI":"10.1007\/s42452-025-07225-5","volume":"7","author":"G Ramesh","year":"2025","unstructured":"Ramesh G, Sahil M, Palan SA et al (2025) A review on NLP zero-shot and few-shot learning: methods and applications. Discov Appl Sci 7:966. https:\/\/doi.org\/10.1007\/s42452-025-07225-5","journal-title":"Discov Appl Sci"}],"container-title":["Neural Computing and Applications"],"original-title":[],"language":"en","link":[{"URL":"https:\/\/link.springer.com\/content\/pdf\/10.1007\/s00521-026-12186-w.pdf","content-type":"application\/pdf","content-version":"vor","intended-application":"text-mining"},{"URL":"https:\/\/link.springer.com\/article\/10.1007\/s00521-026-12186-w","content-type":"text\/html","content-version":"vor","intended-application":"text-mining"},{"URL":"https:\/\/link.springer.com\/content\/pdf\/10.1007\/s00521-026-12186-w.pdf","content-type":"application\/pdf","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2026,7,6]],"date-time":"2026-07-06T05:38:18Z","timestamp":1783316298000},"score":1,"resource":{"primary":{"URL":"https:\/\/link.springer.com\/10.1007\/s00521-026-12186-w"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2026,6]]},"references-count":165,"journal-issue":{"issue":"12","published-print":{"date-parts":[[2026,6]]}},"alternative-id":["12186"],"URL":"https:\/\/doi.org\/10.1007\/s00521-026-12186-w","relation":{},"ISSN":["0941-0643","1433-3058"],"issn-type":[{"value":"0941-0643","type":"print"},{"value":"1433-3058","type":"electronic"}],"subject":[],"published":{"date-parts":[[2026,6]]},"assertion":[{"value":"6 May 2025","order":1,"name":"received","label":"Received","group":{"name":"ArticleHistory","label":"Article History"}},{"value":"11 May 2026","order":2,"name":"accepted","label":"Accepted","group":{"name":"ArticleHistory","label":"Article History"}},{"value":"13 June 2026","order":3,"name":"first_online","label":"First Online","group":{"name":"ArticleHistory","label":"Article History"}},{"order":1,"name":"Ethics","group":{"name":"EthicsHeading","label":"Declarations"}},{"value":"The authors declare that they have no known competing financial interests or personal relationships that could have appeared to influence the work reported in this paper.","order":2,"name":"Ethics","group":{"name":"EthicsHeading","label":"Conflict of interest"}}],"article-number":"485"}}