{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2026,7,18]],"date-time":"2026-07-18T15:22:54Z","timestamp":1784388174568,"version":"3.55.0"},"reference-count":42,"publisher":"MDPI AG","issue":"8","license":[{"start":{"date-parts":[[2025,7,30]],"date-time":"2025-07-30T00:00:00Z","timestamp":1753833600000},"content-version":"vor","delay-in-days":0,"URL":"https:\/\/creativecommons.org\/licenses\/by\/4.0\/"}],"funder":[{"name":"Research on Key Technologies of Robot Vision Perception and Agile Operation for 3C Manufacturing","award":["20232ABC03A09"],"award-info":[{"award-number":["20232ABC03A09"]}]},{"name":"Research on Key Technologies of Robot Vision Perception and Agile Operation for 3C Manufacturing","award":["GJJ180379"],"award-info":[{"award-number":["GJJ180379"]}]},{"name":"Research on Key Technologies of Robot Vision Perception and Agile Operation for 3C Manufacturing","award":["JETRCNGDSS201803"],"award-info":[{"award-number":["JETRCNGDSS201803"]}]},{"name":"Research on the Application of Data Mining in Supporting Decision-Making for the Prediction of Sandstone-Type Uranium Mineralization, Science and Technology Research Project of Jiangxi Provincial Department of Education","award":["20232ABC03A09"],"award-info":[{"award-number":["20232ABC03A09"]}]},{"name":"Research on the Application of Data Mining in Supporting Decision-Making for the Prediction of Sandstone-Type Uranium Mineralization, Science and Technology Research Project of Jiangxi Provincial Department of Education","award":["GJJ180379"],"award-info":[{"award-number":["GJJ180379"]}]},{"name":"Research on the Application of Data Mining in Supporting Decision-Making for the Prediction of Sandstone-Type Uranium Mineralization, Science and Technology Research Project of Jiangxi Provincial Department of Education","award":["JETRCNGDSS201803"],"award-info":[{"award-number":["JETRCNGDSS201803"]}]},{"name":"Research on Machine Learning-based Multi-source Heterogeneous Data Cleaning Techniques for Nuclear Geosciences, Jiangxi Engineering Technology Research Center of Nuclear Geoscience Data Science and System, Open Fund Program","award":["20232ABC03A09"],"award-info":[{"award-number":["20232ABC03A09"]}]},{"name":"Research on Machine Learning-based Multi-source Heterogeneous Data Cleaning Techniques for Nuclear Geosciences, Jiangxi Engineering Technology Research Center of Nuclear Geoscience Data Science and System, Open Fund Program","award":["GJJ180379"],"award-info":[{"award-number":["GJJ180379"]}]},{"name":"Research on Machine Learning-based Multi-source Heterogeneous Data Cleaning Techniques for Nuclear Geosciences, Jiangxi Engineering Technology Research Center of Nuclear Geoscience Data Science and System, Open Fund Program","award":["JETRCNGDSS201803"],"award-info":[{"award-number":["JETRCNGDSS201803"]}]}],"content-domain":{"domain":[],"crossmark-restriction":false},"short-container-title":["Algorithms"],"abstract":"<jats:p>Facial expression recognition (FER) is a critical research direction in artificial intelligence, which is widely used in intelligent interaction, medical diagnosis, security monitoring, and other domains. These applications highlight its considerable practical value and social significance. Face expression recognition models often need to run efficiently on mobile devices or edge devices, so the research on lightweight face expression recognition is particularly important. However, feature extraction and classification methods of lightweight convolutional neural network expression recognition algorithms mostly used at present are not specifically and fully optimized for the characteristics of facial expression images, yet fail to make full use of the feature information in face expression images. To address the lack of facial expression recognition models that are both lightweight and effectively optimized for expression-specific feature extraction, this study proposes a novel network design tailored to the characteristics of facial expressions. In this paper, we refer to the backbone architecture of MobileNet V2 network, and redesign LightExNet, a lightweight convolutional neural network based on the fusion of deep and shallow layers, attention mechanism, and joint loss function, according to the characteristics of the facial expression features. In the network architecture of LightExNet, firstly, deep and shallow features are fused in order to fully extract the shallow features in the original image, reduce the loss of information, alleviate the problem of gradient disappearance when the number of convolutional layers increases, and achieve the effect of multi-scale feature fusion. The MobileNet V2 architecture has also been streamlined to seamlessly integrate deep and shallow networks. Secondly, by combining the own characteristics of face expression features, a new channel and spatial attention mechanism is proposed to obtain the feature information of different expression regions as much as possible for encoding. Thus improve the accuracy of expression recognition effectively. Finally, the improved center loss function is superimposed to further improve the accuracy of face expression classification results, and corresponding measures are taken to significantly reduce the computational volume of the joint loss function. In this paper, LightExNet is tested on the three mainstream face expression datasets: Fer2013, CK+ and RAF-DB, respectively, and the experimental results show that LightExNet has 3.27 M Parameters and 298.27 M Flops, and the accuracy on the three datasets is 69.17%, 97.37%, and 85.97%, respectively. The comprehensive performance of LightExNet is better than the current mainstream lightweight expression recognition algorithms such as MobileNet V2, IE-DBN, Self-Cure Net, Improved MobileViT, MFN, Ada-CM, Parallel CNN(Convolutional Neural Network), etc. Experimental results confirm that LightExNet effectively improves recognition accuracy and computational efficiency while reducing energy consumption and enhancing deployment flexibility. These advantages underscore its strong potential for real-world applications in lightweight facial expression recognition.<\/jats:p>","DOI":"10.3390\/a18080473","type":"journal-article","created":{"date-parts":[[2025,7,30]],"date-time":"2025-07-30T10:55:37Z","timestamp":1753872937000},"page":"473","update-policy":"https:\/\/doi.org\/10.3390\/mdpi_crossmark_policy","source":"Crossref","is-referenced-by-count":10,"title":["A Novel Lightweight Facial Expression Recognition Network Based on Deep Shallow Network Fusion and Attention Mechanism"],"prefix":"10.3390","volume":"18","author":[{"given":"Qiaohe","family":"Yang","sequence":"first","affiliation":[{"name":"Jiangxi Engineering Technology Research Center of Nuclear Geoscience Data Science and System, East China University of Technology, Nanchang 330013, China"},{"name":"School of Artificial Intelligence and Information Engineering, East China University of Technology, Nanchang 330013, China"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Yueshun","family":"He","sequence":"additional","affiliation":[{"name":"Jiangxi Engineering Technology Research Center of Nuclear Geoscience Data Science and System, East China University of Technology, Nanchang 330013, China"},{"name":"School of Artificial Intelligence and Information Engineering, East China University of Technology, Nanchang 330013, China"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Hongmao","family":"Chen","sequence":"additional","affiliation":[{"name":"Jiangxi Engineering Technology Research Center of Nuclear Geoscience Data Science and System, East China University of Technology, Nanchang 330013, China"},{"name":"School of Artificial Intelligence and Information Engineering, East China University of Technology, Nanchang 330013, China"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Youyong","family":"Wu","sequence":"additional","affiliation":[{"name":"Jiangxi Engineering Technology Research Center of Nuclear Geoscience Data Science and System, East China University of Technology, Nanchang 330013, China"},{"name":"School of Artificial Intelligence and Information Engineering, East China University of Technology, Nanchang 330013, China"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Zhihua","family":"Rao","sequence":"additional","affiliation":[{"name":"Jiangxi Engineering Technology Research Center of Nuclear Geoscience Data Science and System, East China University of Technology, Nanchang 330013, China"},{"name":"School of Artificial Intelligence and Information Engineering, East China University of Technology, Nanchang 330013, China"}],"role":[{"vocabulary":"crossref","role":"author"}]}],"member":"1968","published-online":{"date-parts":[[2025,7,30]]},"reference":[{"key":"ref_1","doi-asserted-by":"crossref","first-page":"1195","DOI":"10.1109\/TAFFC.2020.2981446","article-title":"Deep Facial Expression Recognition: A Survey","volume":"13","author":"Li","year":"2022","journal-title":"IEEE Trans. Affect. Comput."},{"key":"ref_2","doi-asserted-by":"crossref","first-page":"124","DOI":"10.1037\/h0030377","article-title":"Constants across cultures in the face and emotion","volume":"17","author":"Ekman","year":"1971","journal-title":"J. Pers. Soc. Psychol."},{"key":"ref_3","doi-asserted-by":"crossref","unstructured":"Liu, W., Zhou, L., and Chen, J. (2021). Face Recognition Based on Lightweight Convolutional Neural Networks. Information, 12.","DOI":"10.3390\/info12050191"},{"key":"ref_4","doi-asserted-by":"crossref","first-page":"1611","DOI":"10.1007\/s11036-019-01366-9","article-title":"Human Behavior Understanding in Big Multimedia Data Using CNN based Facial Expression Recognition","volume":"25","author":"Sajjad","year":"2020","journal-title":"Mob. Netw. Appl."},{"key":"ref_5","doi-asserted-by":"crossref","first-page":"23311","DOI":"10.1007\/s00521-021-06012-8","article-title":"Deep learning-based facial emotion recognition for human\u2013computer interaction applications","volume":"35","author":"Chowdary","year":"2021","journal-title":"Neural Comput. Appl."},{"key":"ref_6","unstructured":"Khaireddin, Y., and Chen, Z. (2021). Facial Emotion Recognition: State of the Art Performance on FER2013. arXiv."},{"key":"ref_7","doi-asserted-by":"crossref","unstructured":"Minaee, S., Minaei, M., and Abdolrashidi, A. (2021). Deep-Emotion: Facial Expression Recognition Using Attentional Convolutional Network. Sensors, 21.","DOI":"10.3390\/s21093046"},{"key":"ref_8","doi-asserted-by":"crossref","first-page":"1","DOI":"10.1016\/j.patrec.2021.03.007","article-title":"Leveraging Recent Advances in Deep Learning for Audio-Visual Emotion Recognition","volume":"146","author":"Schoneveld","year":"2021","journal-title":"Pattern Recognit. Lett."},{"key":"ref_9","doi-asserted-by":"crossref","first-page":"84","DOI":"10.1145\/3065386","article-title":"ImageNet classification with deep convolutional neural networks","volume":"60","author":"Krizhevsky","year":"2012","journal-title":"Commun. ACM"},{"key":"ref_10","unstructured":"Simonyan, K., and Zisserman, A. (2014). Very Deep Convolutional Networks for Large-Scale Image Recognition. arXiv."},{"key":"ref_11","doi-asserted-by":"crossref","unstructured":"Szegedy, C., Liu, W., Jia, Y., Sermanet, P., Reed, S., and Anguelov, D. (2015, January 7\u201312). Going deeper with convolutions. Proceedings of the 2015 IEEE Conference on Computer Vision and Pattern Recognition (CVPR), Boston, MA, USA.","DOI":"10.1109\/CVPR.2015.7298594"},{"key":"ref_12","doi-asserted-by":"crossref","unstructured":"He, K., Zhang, X., Ren, S., and Sun, J. (2016, January 27\u201330). Deep Residual Learning for Image Recognition. Proceedings of the 2016 IEEE Conference on Computer Vision and Pattern Recognition (CVPR), Las Vegas, NV, USA.","DOI":"10.1109\/CVPR.2016.90"},{"key":"ref_13","unstructured":"Iandola, F., Moskewicz, M., Karayev, S., Girshick, R., Darrell, T., and Keutzer, K. (2014). DenseNet: Implementing Efficient ConvNet Descriptor Pyramids. arXiv."},{"key":"ref_14","doi-asserted-by":"crossref","first-page":"784","DOI":"10.21629\/JSEE.2017.04.18","article-title":"Identity-aware convolutional neural networks for facial expression recognition","volume":"28","author":"Chongsheng","year":"2017","journal-title":"J. Syst. Eng. Electron."},{"key":"ref_15","doi-asserted-by":"crossref","unstructured":"Mollahosseini, A., Chan, D., and Mahoor, M.H. (2016, January 7\u201310). Going deeper in facial expression recognition using deep neural networks. Proceedings of the 2016 IEEE Winter Conference on Applications of Computer Vision (WACV), Lake Placid, NY, USA.","DOI":"10.1109\/WACV.2016.7477450"},{"key":"ref_16","doi-asserted-by":"crossref","unstructured":"Liu, K.C., Hsu, C.C., Wang, W.Y., and Chiang, H.H. (2019, January 15\u201318). Facial Expression Recognition Using Merged Convolution Neural Network. Proceedings of the 2019 IEEE 8th Global Conference on Consumer Electronics (GCCE), Las Vegas, NV, USA.","DOI":"10.1109\/GCCE46687.2019.9015479"},{"key":"ref_17","doi-asserted-by":"crossref","unstructured":"Lee, J., Kim, S., Kim, S., Park, J., and Sohn, K. (November, January 27). Context-Aware Emotion Recognition Networks. Proceedings of the 2019 IEEE\/CVF International Conference on Computer Vision (ICCV), Seoul, Republic of Korea.","DOI":"10.1109\/ICCV.2019.01024"},{"key":"ref_18","doi-asserted-by":"crossref","unstructured":"Wang, K., Peng, X., Yang, J., Lu, S., and Qiao, Y. (2020, January 13\u201319). Suppressing Uncertainties for Large-Scale Facial Expression Recognition. Proceedings of the 2020 IEEE\/CVF Conference on Computer Vision and Pattern Recognition (CVPR), Seattle, WA, USA.","DOI":"10.1109\/CVPR42600.2020.00693"},{"key":"ref_19","doi-asserted-by":"crossref","first-page":"898","DOI":"10.1109\/TCDS.2020.3034807","article-title":"Identity\u2013Expression Dual Branch Network for Facial Expression Recognition","volume":"13","author":"Zhang","year":"2021","journal-title":"IEEE Trans. Cogn. Dev. Syst."},{"key":"ref_20","doi-asserted-by":"crossref","first-page":"8094","DOI":"10.1007\/s11227-023-05758-3","article-title":"Three-phases hybrid feature selection for facial expression recognition","volume":"80","author":"Sidhom","year":"2024","journal-title":"J. Supercomput."},{"key":"ref_21","doi-asserted-by":"crossref","first-page":"6499","DOI":"10.1007\/s00521-022-08005-7","article-title":"A deep-learning-based facial expression recognition method using textural features","volume":"35","author":"Mukhopadhyay","year":"2022","journal-title":"Neural Comput. Appl."},{"key":"ref_22","doi-asserted-by":"crossref","first-page":"1057","DOI":"10.1109\/TAFFC.2020.2988264","article-title":"Facial Expression Recognition with Deeply-Supervised Attention Network","volume":"13","author":"Fan","year":"2020","journal-title":"IEEE Trans. Affect. Comput."},{"key":"ref_23","doi-asserted-by":"crossref","unstructured":"Li, H., Wang, N., Yang, X., Wang, X., and Gao, X. (2022, January 18\u201324). Towards Semi-Supervised Deep Facial Expression Recognition with An Adaptive Confidence Margin. Proceedings of the 2022 IEEE\/CVF Conference on Computer Vision and Pattern Recognition (CVPR), New Orleans, LA, USA.","DOI":"10.1109\/CVPR52688.2022.00413"},{"key":"ref_24","unstructured":"Iandola, F.N., Han, S., Moskewicz, M.W., Ashraf, K., Dally, W.J., and Keutzer, K. (2016). SqueezeNet: AlexNet-level accuracy with 50\u00d7 fewer parameters and <0.5MB model size. arXiv."},{"key":"ref_25","unstructured":"Howard, A.G., Zhu, M., Chen, B., Kalenichenko, D., Wang, W., Weyand, T., Andreetto, M., and Adam, H. (2017). MobileNets: Efficient Convolutional Neural Networks for Mobile Vision Applications. arXiv."},{"key":"ref_26","doi-asserted-by":"crossref","unstructured":"Chollet, F. (2017, January 21\u201326). Xception: Deep Learning with Depthwise Separable Convolutions. Proceedings of the 2017 IEEE Conference on Computer Vision and Pattern Recognition (CVPR), Honolulu, HI, USA.","DOI":"10.1109\/CVPR.2017.195"},{"key":"ref_27","doi-asserted-by":"crossref","unstructured":"Sandler, M., Howard, A., Zhu, M., Zhmoginov, A., and Chen, L.-C. (2018, January 18\u201323). MobileNetV2: Inverted Residuals and Linear Bottlenecks. Proceedings of the 2018 IEEE\/CVF Conference on Computer Vision and Pattern Recognition, Salt Lake City, UT, USA.","DOI":"10.1109\/CVPR.2018.00474"},{"key":"ref_28","doi-asserted-by":"crossref","unstructured":"Zhang, X., Zhou, X., Lin, M., and Sun, J. (2018, January 18\u201323). ShuffleNet: An Extremely Efficient Convolutional Neural Network for Mobile Devices. Proceedings of the 2018 IEEE\/CVF Conference on Computer Vision and Pattern Recognition, Salt Lake City, UT, USA.","DOI":"10.1109\/CVPR.2018.00716"},{"key":"ref_29","doi-asserted-by":"crossref","unstructured":"Ma, N., Zhang, X., Zheng, H.-T., and Sun, J. (2018). ShuffleNet V2: Practical Guidelines for Efficient CNN Architecture Design. arXiv.","DOI":"10.1007\/978-3-030-01264-9_8"},{"key":"ref_30","doi-asserted-by":"crossref","unstructured":"Han, K., Wang, Y., Tian, Q., Guo, J., Xu, C., and Xu, C. (2020, January 13\u201319). GhostNet: More Features From Cheap Operations. Proceedings of the 2020 IEEE\/CVF Conference on Computer Vision and Pattern Recognition (CVPR), Seattle, WA, USA.","DOI":"10.1109\/CVPR42600.2020.00165"},{"key":"ref_31","unstructured":"Tan, M., and Le, Q.V. (2019). EfficientNet: Rethinking Model Scaling for Convolutional Neural Networks. arXiv."},{"key":"ref_32","doi-asserted-by":"crossref","unstructured":"Zhu, Q., Zhuang, H., Zhao, M., Xu, S., and Meng, R. (2024). A study on expression recognition based on improved mobilenetV2 network. Sci. Rep., 14.","DOI":"10.1038\/s41598-024-58736-x"},{"key":"ref_33","doi-asserted-by":"crossref","unstructured":"Jiang, B., Li, N., Cui, X., Liu, W., Yu, Z., and Xie, Y. (2024). Research on Facial Expression Recognition Algorithm Based on Lightweight Transformer. Information, 15.","DOI":"10.3390\/info15060321"},{"key":"ref_34","first-page":"71","article-title":"Lightweight Network Based on Multiregion Fusion for Facial Ex-pression Recognition","volume":"60","author":"Hong","year":"2023","journal-title":"Adv. Lasers Optoelectron."},{"key":"ref_35","first-page":"227","article-title":"Expression recognition algorithm for parallel convolutional neural networks","volume":"24","author":"Linlin","year":"2019","journal-title":"Chin. J. Image Graph."},{"key":"ref_36","doi-asserted-by":"crossref","first-page":"4057","DOI":"10.1109\/TIP.2019.2956143","article-title":"Region Attention Networks for Pose and Occlusion Robust Facial Expression Recognition","volume":"29","author":"Wang","year":"2020","journal-title":"IEEE Trans. Image Process."},{"key":"ref_37","doi-asserted-by":"crossref","first-page":"110951","DOI":"10.1016\/j.patcog.2024.110951","article-title":"POSTER++: A Simpler and Stronger Facial Expression Recognition Network","volume":"157","author":"Mao","year":"2025","journal-title":"Pattern Recognit."},{"key":"ref_38","doi-asserted-by":"crossref","first-page":"593","DOI":"10.1007\/s41095-023-0369-x","article-title":"CF-DAN: Facial-expression Recognition Based on Cross-Fusion Dual-Attention Network","volume":"10","author":"Zhang","year":"2024","journal-title":"Comput. Vis. Media"},{"key":"ref_39","doi-asserted-by":"crossref","unstructured":"Howard, A., Sandler, M., Chu, G., Chen, L.-C., Chen, B., Tan, M., Wang, W., Zhu, Y., Pang, R., and Vasudevan, V. (November, January 27). Searching for MobileNetV3. Proceedings of the 2019 IEEE\/CVF International Conference on Computer Vision (ICCV), Seoul, Republic of Korea.","DOI":"10.1109\/ICCV.2019.00140"},{"key":"ref_40","doi-asserted-by":"crossref","unstructured":"Woo, S., Park, J., Lee, J.-Y., and Kweon, I.S. (2018). CBAM: Convolutional Block Attention Module, Springer.","DOI":"10.1007\/978-3-030-01234-2_1"},{"key":"ref_41","doi-asserted-by":"crossref","unstructured":"Lucey, P., Cohn, J.F., Kanade, T., Saragih, J., Ambadar, Z., and Matthews, I. (2010, January 13\u201318). The extended cohn-kanade dataset (ck+): A complete dataset for action unit and emotion-specified expression. Proceedings of the 2010 IEEE Computer Society Conference on Computer Vision and Pattern Recognition-workshops, San Francisco, CA, USA.","DOI":"10.1109\/CVPRW.2010.5543262"},{"key":"ref_42","doi-asserted-by":"crossref","unstructured":"Li, S., Deng, W., and Du, J. (2017, January 21\u201326). Reliable crowdsourcing and deep locality-preserving learning for expression recognition in the wild. Proceedings of the 2017 IEEE Conference on Computer Vision and Pattern Recognition (CVPR), Honolulu, HI, USA.","DOI":"10.1109\/CVPR.2017.277"}],"container-title":["Algorithms"],"original-title":[],"language":"en","link":[{"URL":"https:\/\/www.mdpi.com\/1999-4893\/18\/8\/473\/pdf","content-type":"unspecified","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2025,10,9]],"date-time":"2025-10-09T18:18:40Z","timestamp":1760033920000},"score":1,"resource":{"primary":{"URL":"https:\/\/www.mdpi.com\/1999-4893\/18\/8\/473"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2025,7,30]]},"references-count":42,"journal-issue":{"issue":"8","published-online":{"date-parts":[[2025,8]]}},"alternative-id":["a18080473"],"URL":"https:\/\/doi.org\/10.3390\/a18080473","relation":{},"ISSN":["1999-4893"],"issn-type":[{"value":"1999-4893","type":"electronic"}],"subject":[],"published":{"date-parts":[[2025,7,30]]}}}