{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2026,6,30]],"date-time":"2026-06-30T15:39:21Z","timestamp":1782833961510,"version":"3.54.5"},"reference-count":40,"publisher":"Association for Computing Machinery (ACM)","issue":"3","license":[{"start":{"date-parts":[[2023,4,19]],"date-time":"2023-04-19T00:00:00Z","timestamp":1681862400000},"content-version":"vor","delay-in-days":0,"URL":"https:\/\/www.acm.org\/publications\/policies\/copyright_policy#Background"}],"funder":[{"DOI":"10.13039\/501100001809","name":"National Natural Science Foundation of China","doi-asserted-by":"crossref","award":["62172438"],"award-info":[{"award-number":["62172438"]}],"id":[{"id":"10.13039\/501100001809","id-type":"DOI","asserted-by":"crossref"}]},{"name":"Key Project of Shenzhen City Special Fund for Fundamental Research","award":["202208183000751"],"award-info":[{"award-number":["202208183000751"]}]},{"DOI":"10.13039\/501100012166","name":"National Key R&D Program of China","doi-asserted-by":"crossref","award":["2021YFB3900601"],"award-info":[{"award-number":["2021YFB3900601"]}],"id":[{"id":"10.13039\/501100012166","id-type":"DOI","asserted-by":"crossref"}]},{"DOI":"10.13039\/501100012226","name":"Fundamental Research Funds for the Central Universities","doi-asserted-by":"crossref","award":["B220202074"],"award-info":[{"award-number":["B220202074"]}],"id":[{"id":"10.13039\/501100012226","id-type":"DOI","asserted-by":"crossref"}]},{"name":"Fundamental Research Funds for the Central Universities, JLU, Joint Foundation of the Ministry of Education","award":["8091B022123"],"award-info":[{"award-number":["8091B022123"]}]},{"name":"Kempe Foundation, Sweden"}],"content-domain":{"domain":["dl.acm.org"],"crossmark-restriction":true},"short-container-title":["ACM Trans. Embed. Comput. Syst."],"published-print":{"date-parts":[[2023,5,31]]},"abstract":"<jats:p>Facial Expression Recognition (FER) in the wild poses significant challenges due to realistic occlusions, illumination, scale, and head pose variations of the facial images. In this article, we propose an Edge-AI-driven framework for FER. On the algorithms aspect, we propose two attention modules, Arbitrary-oriented Spatial Pooling (ASP) and Scalable Frequency Pooling (SFP), for effective feature extraction to improve classification accuracy. On the systems aspect, we propose an edge-cloud joint inference architecture for FER to achieve low-latency inference, consisting of a lightweight backbone network running on the edge device, and two optional attention modules partially offloaded to the cloud. Performance evaluation demonstrates that our approach achieves a good balance between classification accuracy and inference latency.<\/jats:p>","DOI":"10.1145\/3587038","type":"journal-article","created":{"date-parts":[[2023,3,6]],"date-time":"2023-03-06T12:38:38Z","timestamp":1678106318000},"page":"1-17","update-policy":"https:\/\/doi.org\/10.1145\/crossmark-policy","source":"Crossref","is-referenced-by-count":66,"title":["Edge-AI-Driven Framework with Efficient Mobile Network Design for Facial Expression Recognition"],"prefix":"10.1145","volume":"22","author":[{"ORCID":"https:\/\/orcid.org\/0000-0003-3022-3718","authenticated-orcid":false,"given":"Yirui","family":"Wu","sequence":"first","affiliation":[{"name":"Hohai University, Jiangsu Province, China"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"ORCID":"https:\/\/orcid.org\/0000-0003-4324-4016","authenticated-orcid":false,"given":"Lilai","family":"Zhang","sequence":"additional","affiliation":[{"name":"Hohai University, Jiangsu Province, China"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"ORCID":"https:\/\/orcid.org\/0000-0003-4228-2774","authenticated-orcid":false,"given":"Zonghua","family":"Gu","sequence":"additional","affiliation":[{"name":"Ume\u00e5 University, Ume\u00e5, Sweden"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"ORCID":"https:\/\/orcid.org\/0000-0003-0350-4055","authenticated-orcid":false,"given":"Hu","family":"Lu","sequence":"additional","affiliation":[{"name":"Jiangsu University, Jiangsu Province, China"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"ORCID":"https:\/\/orcid.org\/0000-0001-7013-9081","authenticated-orcid":false,"given":"Shaohua","family":"Wan","sequence":"additional","affiliation":[{"name":"University of Electronic Science and Technology of China, Guangdong Province, China"}],"role":[{"vocabulary":"crossref","role":"author"}]}],"member":"320","published-online":{"date-parts":[[2023,4,19]]},"reference":[{"key":"e_1_3_2_2_2","doi-asserted-by":"crossref","first-page":"630","DOI":"10.1109\/ASPDAC.2015.7059079","volume-title":"Proceedings of the 20th Asia and South Pacific Design Automation Conference","author":"Al-bayati Zaid","year":"2015","unstructured":"Zaid Al-bayati, Qingling Zhao, Ahmed Youssef, Haibo Zeng, and Zonghua Gu. 2015. Enhanced partitioned scheduling of mixed-criticality systems on multicore platforms. In Proceedings of the 20th Asia and South Pacific Design Automation Conference. IEEE, 630\u2013635."},{"key":"e_1_3_2_3_2","first-page":"279","volume-title":"Proceedings of the ACM International Conference on Multimodal Interaction","author":"Barsoum Emad","year":"2016","unstructured":"Emad Barsoum, Cha Zhang, Cristian Canton Ferrer, and Zhengyou Zhang. 2016. Training deep networks for facial expression recognition with crowd-sourced label distribution. In Proceedings of the ACM International Conference on Multimodal Interaction. 279\u2013283."},{"key":"e_1_3_2_4_2","first-page":"929","volume-title":"Proceedings of the 35th AAAI Conference on Artificial Intelligence","author":"Behera Ardhendu","year":"2021","unstructured":"Ardhendu Behera, Zachary Wharton, Pradeep R. P. G. Hewage, and Asish Bera. 2021. Context-aware attentional pooling (CAP) for fine-grained visual classification. In Proceedings of the 35th AAAI Conference on Artificial Intelligence. 929\u2013937."},{"issue":"5","key":"e_1_3_2_5_2","first-page":"82:1\u201382:22","article-title":"Memory- and communication-aware model compression for distributed deep learning inference on IoT","volume":"18","author":"Bhardwaj Kartikeya","year":"2019","unstructured":"Kartikeya Bhardwaj, Chingyi Lin, Anderson L. Sartor, and Radu Marculescu. 2019. Memory- and communication-aware model compression for distributed deep learning inference on IoT. ACM Transactions on Embedded Computing Systems 18, 5s (2019), 82:1\u201382:22.","journal-title":"ACM Transactions on Embedded Computing Systems"},{"key":"e_1_3_2_6_2","doi-asserted-by":"publisher","DOI":"10.1007\/s10462-020-09816-7"},{"key":"e_1_3_2_7_2","first-page":"2106","volume-title":"Proceedings of the IEEE International Conference on Computer Vision Workshops","author":"Dhall Abhinav","year":"2011","unstructured":"Abhinav Dhall, Roland Goecke, Simon Lucey, and Tom Gedeon. 2011. Static facial expression analysis in tough conditions: Data, evaluation protocol, and benchmark. In Proceedings of the IEEE International Conference on Computer Vision Workshops. 2106\u20132112."},{"key":"e_1_3_2_8_2","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR.2016.90"},{"key":"e_1_3_2_9_2","first-page":"13713","volume-title":"Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition","author":"Hou Qibin","year":"2021","unstructured":"Qibin Hou, Daquan Zhou, and Jiashi Feng. 2021. Coordinate attention for efficient mobile network design. In Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition. 13713\u201313722."},{"key":"e_1_3_2_10_2","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR.2018.00745"},{"issue":"1","key":"e_1_3_2_11_2","first-page":"4:1\u20134:24","article-title":"Design and optimization of energy-accuracy tradeoff networks for mobile platforms via pretrained deep models","volume":"19","author":"Jayakodi Nitthilan Kanappan","year":"2020","unstructured":"Nitthilan Kanappan Jayakodi, Syrine Belakaria, Aryan Deshwal, and Janardhan Rao Doppa. 2020. Design and optimization of energy-accuracy tradeoff networks for mobile platforms via pretrained deep models. ACM Transactions on Embedded Computing Systems 19, 1 (2020), 4:1\u20134:24.","journal-title":"ACM Transactions on Embedded Computing Systems"},{"key":"e_1_3_2_12_2","first-page":"23:1\u201323:9","volume-title":"Proceedings of the IEEE\/ACM International Conference On Computer Aided Design","author":"Jayakodi Nitthilan Kanappan","year":"2020","unstructured":"Nitthilan Kanappan Jayakodi, Janardhan Rao Doppa, and Partha Pratim Pande. 2020. SETGAN: Scale and energy tradeoff GANs for image applications on mobile platforms. In Proceedings of the IEEE\/ACM International Conference On Computer Aided Design. 23:1\u201323:9."},{"key":"e_1_3_2_13_2","doi-asserted-by":"publisher","DOI":"10.1016\/j.patcog.2021.108159"},{"key":"e_1_3_2_14_2","doi-asserted-by":"publisher","DOI":"10.1080\/02699930903485076"},{"key":"e_1_3_2_15_2","doi-asserted-by":"publisher","DOI":"10.1109\/TWC.2019.2946140"},{"key":"e_1_3_2_16_2","first-page":"2209","volume-title":"Proceedings of the International Conference on Pattern Recognition","author":"Li Yong","year":"2018","unstructured":"Yong Li, Jiabei Zeng, Shiguang Shan, and Xilin Chen. 2018. Patch-gated CNN for occlusion-aware facial expression recognition. In Proceedings of the International Conference on Pattern Recognition. 2209\u20132214."},{"key":"e_1_3_2_17_2","doi-asserted-by":"publisher","DOI":"10.1109\/TIP.2018.2886767"},{"key":"e_1_3_2_18_2","first-page":"1","volume-title":"Proceedings of the 2014 IEEE 20th International Conference on Embedded and Real-Time Computing Systems and Applications","author":"Liu Guangdong","year":"2014","unstructured":"Guangdong Liu, Ying Lu, Shige Wang, and Zonghua Gu. 2014. Partitioned multiprocessor scheduling of mixed-criticality parallel jobs. In Proceedings of the 2014 IEEE 20th International Conference on Embedded and Real-Time Computing Systems and Applications. IEEE, 1\u201310."},{"key":"e_1_3_2_19_2","doi-asserted-by":"publisher","DOI":"10.1016\/j.patcog.2018.07.016"},{"key":"e_1_3_2_20_2","doi-asserted-by":"crossref","first-page":"e7351","DOI":"10.1002\/cpe.7351","article-title":"LRP-based network pruning and policy distillation of robust and non-robust DRL agents for embedded systems","author":"Luan Siyu","year":"2023","unstructured":"Siyu Luan, Zonghua Gu, Rui Xu, Qingling Zhao, and Gang Chen. 2023. LRP-based network pruning and policy distillation of robust and non-robust DRL agents for embedded systems. Concurrency and Computation: Practice and Experience (2023), e7351.","journal-title":"Concurrency and Computation: Practice and Experience"},{"key":"e_1_3_2_21_2","first-page":"122","volume-title":"Proceedings of the European Conference on Computer Vision","author":"Ma Ningning","year":"2018","unstructured":"Ningning Ma, Xiangyu Zhang, Hai-Tao Zheng, and Jian Sun. 2018. ShuffleNet V2: Practical guidelines for efficient CNN architecture design. In Proceedings of the European Conference on Computer Vision. 122\u2013138."},{"key":"e_1_3_2_22_2","unstructured":"Wenjia Meng Zonghua Gu Ming Zhang and Zhaohui Wu. 2017. Two-bit networks for deep learning on resource-constrained embedded devices. CoRR abs\/1701.00485 (2017)."},{"key":"e_1_3_2_23_2","first-page":"6913","volume-title":"Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition","author":"Oh Youngmin","year":"2021","unstructured":"Youngmin Oh, Beomjun Kim, and Bumsub Ham. 2021. Background-aware pooling and noise-aware loss for weakly-supervised semantic segmentation. In Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition. 6913\u20136922."},{"key":"e_1_3_2_24_2","first-page":"475","volume-title":"Proceedings of the Design, Automation, and Test in Europe Conference and Exhibition","author":"Panda Priyadarshini","year":"2016","unstructured":"Priyadarshini Panda, Abhronil Sengupta, and Kaushik Roy. 2016. Conditional deep learning for energy-efficient and enhanced pattern recognition. In Proceedings of the Design, Automation, and Test in Europe Conference and Exhibition. 475\u2013480."},{"key":"e_1_3_2_25_2","article-title":"Bam: Bottleneck attention module","author":"Park Jongchan","year":"2018","unstructured":"Jongchan Park, Sanghyun Woo, Joon-Young Lee, and In So Kweon. 2018. Bam: Bottleneck attention module. In Proceedings of British Machine Vision Conference. 147\u2013160.","journal-title":"Proceedings of British Machine Vision Conference"},{"key":"e_1_3_2_26_2","doi-asserted-by":"publisher","DOI":"10.1109\/ICCV48922.2021.00082"},{"key":"e_1_3_2_27_2","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR.2018.00474"},{"key":"e_1_3_2_28_2","unstructured":"Mathijs Schuurmans Maxim Berman and Matthew B. Blaschko. 2018. Efficient semantic image segmentation with superpixel pooling. CoRR abs\/1806.02705 (2018)."},{"key":"e_1_3_2_29_2","unstructured":"Vladislav Sovrasov. 2022. Ptflops: A flops counting tool for neural networks in pytorch framework. https:\/\/github.com\/sovrasov\/flops-counter.pytorch."},{"key":"e_1_3_2_30_2","first-page":"23","volume-title":"Proceedings of the International Conference on Computer-Aided Design","author":"Stamoulis Dimitrios","year":"2018","unstructured":"Dimitrios Stamoulis, Ting-Wu (Rudy) Chin, Anand Krishnan Prakash, Haocheng Fang, Sribhuvan Sajja, Mitchell Bognar, and Diana Marculescu. 2018. Designing adaptive neural networks for energy-constrained image classification. In Proceedings of the International Conference on Computer-Aided Design. 23."},{"key":"e_1_3_2_31_2","doi-asserted-by":"publisher","DOI":"10.1109\/MCSE.2011.37"},{"key":"e_1_3_2_32_2","first-page":"1674","volume-title":"Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition","author":"Vesdapunt Noranart","year":"2021","unstructured":"Noranart Vesdapunt and Baoyuan Wang. 2021. CRFace: Confidence ranker for model-agnostic face detection refinement. In Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition. 1674\u20131684."},{"key":"e_1_3_2_33_2","doi-asserted-by":"publisher","DOI":"10.1109\/TIP.2019.2956143"},{"key":"e_1_3_2_34_2","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR42600.2020.01155"},{"key":"e_1_3_2_35_2","first-page":"16195","volume-title":"Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition","author":"Wang Wenjing","year":"2021","unstructured":"Wenjing Wang, Wenhan Yang, and Jiaying Liu. 2021. HLA-Face: Joint high-low adaptation for low light face detection. In Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition. 16195\u201316204."},{"key":"e_1_3_2_36_2","first-page":"3","volume-title":"Proceedings of the European Conference on Computer Vision","author":"Woo Sanghyun","year":"2018","unstructured":"Sanghyun Woo, Jongchan Park, Joon-Young Lee, and In So Kweon. 2018. CBAM: Convolutional block attention module. In Proceedings of the European Conference on Computer Vision. Vol. 11211. 3\u201319."},{"key":"e_1_3_2_37_2","doi-asserted-by":"publisher","DOI":"10.1109\/TNSE.2022.3151502"},{"key":"e_1_3_2_38_2","doi-asserted-by":"publisher","DOI":"10.1016\/j.patcog.2019.03.019"},{"key":"e_1_3_2_39_2","first-page":"4003","volume-title":"Proceedings of the Conference on Computer Vision and Pattern Recognition","author":"Zhai Shuangfei","year":"2017","unstructured":"Shuangfei Zhai, Hui Wu, Abhishek Kumar, Yu Cheng, Yongxi Lu, Zhongfei Zhang, and Rog\u00e9rio Schmidt Feris. 2017. S3Pool: Pooling with stochastic spatial sampling. In Proceedings of the Conference on Computer Vision and Pattern Recognition. 4003\u20134011."},{"key":"e_1_3_2_40_2","doi-asserted-by":"publisher","DOI":"10.1109\/TIP.2021.3093397"},{"key":"e_1_3_2_41_2","first-page":"3510","volume-title":"Proceedings of the AAAI","author":"Zhao Zengqun","year":"2021","unstructured":"Zengqun Zhao, Qingshan Liu, and Feng Zhou. 2021. Robust lightweight facial expression recognition network with label distribution training. In Proceedings of the AAAI. 3510\u20133519."}],"container-title":["ACM Transactions on Embedded Computing Systems"],"original-title":[],"language":"en","link":[{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/3587038","content-type":"unspecified","content-version":"vor","intended-application":"text-mining"},{"URL":"https:\/\/dl.acm.org\/doi\/pdf\/10.1145\/3587038","content-type":"unspecified","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2025,6,17]],"date-time":"2025-06-17T16:37:33Z","timestamp":1750178253000},"score":1,"resource":{"primary":{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/3587038"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2023,4,19]]},"references-count":40,"journal-issue":{"issue":"3","published-print":{"date-parts":[[2023,5,31]]}},"alternative-id":["10.1145\/3587038"],"URL":"https:\/\/doi.org\/10.1145\/3587038","relation":{},"ISSN":["1539-9087","1558-3465"],"issn-type":[{"value":"1539-9087","type":"print"},{"value":"1558-3465","type":"electronic"}],"subject":[],"published":{"date-parts":[[2023,4,19]]},"assertion":[{"value":"2022-05-06","order":0,"name":"received","label":"Received","group":{"name":"publication_history","label":"Publication History"}},{"value":"2023-02-17","order":1,"name":"accepted","label":"Accepted","group":{"name":"publication_history","label":"Publication History"}},{"value":"2023-04-19","order":2,"name":"published","label":"Published","group":{"name":"publication_history","label":"Publication History"}}]}}