{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2026,6,16]],"date-time":"2026-06-16T04:59:00Z","timestamp":1781585940218,"version":"3.54.5"},"publisher-location":"New York, NY, USA","reference-count":52,"publisher":"ACM","license":[{"start":{"date-parts":[[2022,10,10]],"date-time":"2022-10-10T00:00:00Z","timestamp":1665360000000},"content-version":"vor","delay-in-days":0,"URL":"https:\/\/www.acm.org\/publications\/policies\/copyright_policy#Background"}],"content-domain":{"domain":["dl.acm.org"],"crossmark-restriction":true},"short-container-title":[],"published-print":{"date-parts":[[2022,10,10]]},"DOI":"10.1145\/3503161.3548407","type":"proceedings-article","created":{"date-parts":[[2022,10,10]],"date-time":"2022-10-10T15:43:01Z","timestamp":1665416581000},"page":"33-41","update-policy":"https:\/\/doi.org\/10.1145\/crossmark-policy","source":"Crossref","is-referenced-by-count":16,"title":["SER30K"],"prefix":"10.1145","author":[{"given":"Shengzhe","family":"Liu","sequence":"first","affiliation":[{"name":"Nankai University, Tianjin, China"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Xin","family":"Zhang","sequence":"additional","affiliation":[{"name":"Nankai University, Tianjin, China"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Jufeng","family":"Yang","sequence":"additional","affiliation":[{"name":"Nankai University, Tianjin, China"}],"role":[{"vocabulary":"crossref","role":"author"}]}],"member":"320","published-online":{"date-parts":[[2022,10,10]]},"reference":[{"key":"e_1_3_2_2_1_1","volume-title":"Artemis: Affective language for visual art. In CVPR.","author":"Achlioptas Panos","year":"2021","unstructured":"Panos Achlioptas , Maks Ovsjanikov , Kilichbek Haydarov , Mohamed Elhoseiny , and Leonidas J Guibas . 2021 . Artemis: Affective language for visual art. In CVPR. Panos Achlioptas, Maks Ovsjanikov, Kilichbek Haydarov, Mohamed Elhoseiny, and Leonidas J Guibas. 2021. Artemis: Affective language for visual art. In CVPR."},{"key":"e_1_3_2_2_2_1","volume-title":"Jamie Ryan Kiros, and Geoffrey E Hinton","author":"Ba Jimmy Lei","year":"2016","unstructured":"Jimmy Lei Ba , Jamie Ryan Kiros, and Geoffrey E Hinton . 2016 . Layer normalization. arXiv preprint arXiv:1607.06450 (2016). Jimmy Lei Ba, Jamie Ryan Kiros, and Geoffrey E Hinton. 2016. Layer normalization. arXiv preprint arXiv:1607.06450 (2016)."},{"key":"e_1_3_2_2_3_1","doi-asserted-by":"crossref","unstructured":"Damian Borth Rongrong Ji Tao Chen Thomas Breuel and Shih-Fu Chang. 2013. Large-scale visual sentiment ontology and detectors using adjective noun pairs. In ACM MM.  Damian Borth Rongrong Ji Tao Chen Thomas Breuel and Shih-Fu Chang. 2013. Large-scale visual sentiment ontology and detectors using adjective noun pairs. In ACM MM.","DOI":"10.1145\/2502081.2502282"},{"key":"e_1_3_2_2_4_1","volume-title":"IEMOCAP: Interactive emotional dyadic motion capture database. Language resources and evaluation","author":"Busso Carlos","year":"2008","unstructured":"Carlos Busso , Murtaza Bulut , Chi-Chun Lee , Abe Kazemzadeh , Emily Mower , Samuel Kim , Jeannette N Chang , Sungbok Lee , and Shrikanth S Narayanan . 2008 . IEMOCAP: Interactive emotional dyadic motion capture database. Language resources and evaluation , Vol. 42 , 4 (2008), 335--359. Carlos Busso, Murtaza Bulut, Chi-Chun Lee, Abe Kazemzadeh, Emily Mower, Samuel Kim, Jeannette N Chang, Sungbok Lee, and Shrikanth S Narayanan. 2008. IEMOCAP: Interactive emotional dyadic motion capture database. Language resources and evaluation, Vol. 42, 4 (2008), 335--359."},{"key":"e_1_3_2_2_5_1","doi-asserted-by":"crossref","unstructured":"Yitao Cai Huiyu Cai and Xiaojun Wan. 2019. Multi-modal sarcasm detection in twitter with hierarchical fusion model. In ACL.  Yitao Cai Huiyu Cai and Xiaojun Wan. 2019. Multi-modal sarcasm detection in twitter with hierarchical fusion model. In ACL.","DOI":"10.18653\/v1\/P19-1239"},{"key":"e_1_3_2_2_6_1","volume-title":"Deepsentibank: Visual sentiment concept classification with deep convolutional neural networks. arXiv preprint arXiv:1410.(2014).","author":"Chen Tao","year":"2014","unstructured":"Tao Chen , Damian Borth , Trevor Darrell , and Shih-Fu Chang . 2014 . Deepsentibank: Visual sentiment concept classification with deep convolutional neural networks. arXiv preprint arXiv:1410.(2014). Tao Chen, Damian Borth, Trevor Darrell, and Shih-Fu Chang. 2014. Deepsentibank: Visual sentiment concept classification with deep convolutional neural networks. arXiv preprint arXiv:1410.(2014)."},{"key":"e_1_3_2_2_7_1","volume-title":"Imagenet: A large-scale hierarchical image database. In CVPR.","author":"Deng Jia","year":"2009","unstructured":"Jia Deng , Wei Dong , Richard Socher , Li-Jia Li , Kai Li , and Li Fei-Fei . 2009 . Imagenet: A large-scale hierarchical image database. In CVPR. Jia Deng, Wei Dong, Richard Socher, Li-Jia Li, Kai Li, and Li Fei-Fei. 2009. Imagenet: A large-scale hierarchical image database. In CVPR."},{"key":"e_1_3_2_2_8_1","volume-title":"Bert: Pre-training of deep bidirectional transformers for language understanding. arXiv preprint arXiv:1810.04805","author":"Devlin Jacob","year":"2018","unstructured":"Jacob Devlin , Ming-Wei Chang , Kenton Lee , and Kristina Toutanova . 2018 . Bert: Pre-training of deep bidirectional transformers for language understanding. arXiv preprint arXiv:1810.04805 (2018). Jacob Devlin, Ming-Wei Chang, Kenton Lee, and Kristina Toutanova. 2018. Bert: Pre-training of deep bidirectional transformers for language understanding. arXiv preprint arXiv:1810.04805 (2018)."},{"key":"e_1_3_2_2_9_1","unstructured":"Alexey Dosovitskiy Lucas Beyer Alexander Kolesnikov Dirk Weissenborn Xiaohua Zhai Thomas Unterthiner Mostafa Dehghani Matthias Minderer Georg Heigold Sylvain Gelly etal 2020. An image is worth 16x16 words: Transformers for image recognition at scale. arXiv preprint arXiv:2010.11929 (2020).  Alexey Dosovitskiy Lucas Beyer Alexander Kolesnikov Dirk Weissenborn Xiaohua Zhai Thomas Unterthiner Mostafa Dehghani Matthias Minderer Georg Heigold Sylvain Gelly et al. 2020. An image is worth 16x16 words: Transformers for image recognition at scale. arXiv preprint arXiv:2010.11929 (2020)."},{"key":"e_1_3_2_2_10_1","volume-title":"Gated attention fusion network for multimodal sentiment classification. Knowledge-Based Systems","author":"Du Yongping","year":"2022","unstructured":"Yongping Du , Yang Liu , Zhi Peng , and Xingnan Jin . 2022. Gated attention fusion network for multimodal sentiment classification. Knowledge-Based Systems ( 2022 ), 108107. Yongping Du, Yang Liu, Zhi Peng, and Xingnan Jin. 2022. Gated attention fusion network for multimodal sentiment classification. Knowledge-Based Systems (2022), 108107."},{"key":"e_1_3_2_2_11_1","volume-title":"An argument for basic emotions. Cognition & emotion","author":"Ekman Paul","year":"1992","unstructured":"Paul Ekman . 1992. An argument for basic emotions. Cognition & emotion ( 1992 ). Paul Ekman. 1992. An argument for basic emotions. Cognition & emotion (1992)."},{"key":"e_1_3_2_2_12_1","volume-title":"Towards expressive communication with internet memes: A new multimodal conversation dataset and benchmark. arXiv preprint arXiv:2109.01839","author":"Fei Zhengcong","year":"2021","unstructured":"Zhengcong Fei , Zekang Li , Jinchao Zhang , Yang Feng , and Jie Zhou . 2021. Towards expressive communication with internet memes: A new multimodal conversation dataset and benchmark. arXiv preprint arXiv:2109.01839 ( 2021 ). Zhengcong Fei, Zekang Li, Jinchao Zhang, Yang Feng, and Jie Zhou. 2021. Towards expressive communication with internet memes: A new multimodal conversation dataset and benchmark. arXiv preprint arXiv:2109.01839 (2021)."},{"key":"e_1_3_2_2_13_1","doi-asserted-by":"crossref","unstructured":"Yang Gao Oscar Beijbom Ning Zhang and Trevor Darrell. 2016. Compact bilinear pooling. In CVPR.  Yang Gao Oscar Beijbom Ning Zhang and Trevor Darrell. 2016. Compact bilinear pooling. In CVPR.","DOI":"10.1109\/CVPR.2016.41"},{"key":"e_1_3_2_2_14_1","volume-title":"Transfg: A transformer architecture for fine-grained recognition. arXiv preprint arXiv:2103.07976","author":"He Ju","year":"2021","unstructured":"Ju He , Jie-Neng Chen , Shuai Liu , Adam Kortylewski , Cheng Yang , Yutong Bai , Changhu Wang , and Alan Yuille . 2021 . Transfg: A transformer architecture for fine-grained recognition. arXiv preprint arXiv:2103.07976 (2021). Ju He, Jie-Neng Chen, Shuai Liu, Adam Kortylewski, Cheng Yang, Yutong Bai, Changhu Wang, and Alan Yuille. 2021. Transfg: A transformer architecture for fine-grained recognition. arXiv preprint arXiv:2103.07976 (2021)."},{"key":"e_1_3_2_2_15_1","unstructured":"Kaiming He Xiangyu Zhang Shaoqing Ren and Jian Sun. 2016. Deep residual learning for image recognition. In CVPR.  Kaiming He Xiangyu Zhang Shaoqing Ren and Jian Sun. 2016. Deep residual learning for image recognition. In CVPR."},{"key":"e_1_3_2_2_16_1","unstructured":"Xiaohao He Huijun Zhang Ningyun Li Ling Feng and Feng Zheng. 2019. A multi-attentive pyramidal model for visual sentiment analysis. In IJCNN.  Xiaohao He Huijun Zhang Ningyun Li Ling Feng and Feng Zheng. 2019. A multi-attentive pyramidal model for visual sentiment analysis. In IJCNN."},{"key":"e_1_3_2_2_17_1","doi-asserted-by":"crossref","unstructured":"Susan Herring and Ashley Dainas. 2017. \"Nice picture comment!\" Graphicons in Facebook comment threads. In HICSS.  Susan Herring and Ashley Dainas. 2017. \"Nice picture comment!\" Graphicons in Facebook comment threads. In HICSS.","DOI":"10.24251\/HICSS.2017.264"},{"key":"e_1_3_2_2_18_1","volume-title":"Sticker and emoji use in Facebook messenger: Implications for graphicon change. Journal of Computer-Mediated Communication","author":"Konrad Artie","year":"2020","unstructured":"Artie Konrad , Susan C Herring , and David Choi . 2020. Sticker and emoji use in Facebook messenger: Implications for graphicon change. Journal of Computer-Mediated Communication ( 2020 ). Artie Konrad, Susan C Herring, and David Choi. 2020. Sticker and emoji use in Facebook messenger: Implications for graphicon change. Journal of Computer-Mediated Communication (2020)."},{"key":"e_1_3_2_2_19_1","first-page":"2755","article-title":"Context based emotion recognition using emotic dataset","volume":"42","author":"Kosti Ronak","year":"2019","unstructured":"Ronak Kosti , Jose M Alvarez , Adria Recasens , and Agata Lapedriza . 2019 . Context based emotion recognition using emotic dataset . IEEE Transactions on Pattern Analysis and Machine Intelligence , Vol. 42 , 11 (2019), 2755 -- 2766 . Ronak Kosti, Jose M Alvarez, Adria Recasens, and Agata Lapedriza. 2019. Context based emotion recognition using emotic dataset. IEEE Transactions on Pattern Analysis and Machine Intelligence, Vol. 42, 11 (2019), 2755--2766.","journal-title":"IEEE Transactions on Pattern Analysis and Machine Intelligence"},{"key":"e_1_3_2_2_20_1","volume-title":"Imagenet classification with deep convolutional neural networks. NeurIPS","author":"Krizhevsky Alex","year":"2012","unstructured":"Alex Krizhevsky , Ilya Sutskever , and Geoffrey E Hinton . 2012. Imagenet classification with deep convolutional neural networks. NeurIPS ( 2012 ). Alex Krizhevsky, Ilya Sutskever, and Geoffrey E Hinton. 2012. Imagenet classification with deep convolutional neural networks. NeurIPS (2012)."},{"key":"e_1_3_2_2_21_1","unstructured":"Peter J Lang Margaret M Bradley Bruce N Cuthbert etal 1997. International affective picture system (IAPS): Technical manual and affective ratings. NIMH Center for the Study of Emotion and Attention (1997).  Peter J Lang Margaret M Bradley Bruce N Cuthbert et al. 1997. International affective picture system (IAPS): Technical manual and affective ratings. NIMH Center for the Study of Emotion and Attention (1997)."},{"key":"e_1_3_2_2_22_1","doi-asserted-by":"crossref","unstructured":"Jun Ling Han Xue Li Song Shuhui Yang Rong Xie and Xiao Gu. 2020. Toward fine-grained facial expression manipulation. In ECCV.  Jun Ling Han Xue Li Song Shuhui Yang Rong Xie and Xiao Gu. 2020. Toward fine-grained facial expression manipulation. In ECCV.","DOI":"10.1007\/978-3-030-58604-1_3"},{"key":"e_1_3_2_2_23_1","doi-asserted-by":"crossref","unstructured":"Rameswar Panda Jianming Zhang Haoxiang Li Joon-Young Lee Xin Lu and Amit K Roy-Chowdhury. 2018. Contemplating visual emotions: Understanding and overcoming dataset bias. In ECCV.  Rameswar Panda Jianming Zhang Haoxiang Li Joon-Young Lee Xin Lu and Amit K Roy-Chowdhury. 2018. Contemplating visual emotions: Understanding and overcoming dataset bias. In ECCV.","DOI":"10.1007\/978-3-030-01216-8_36"},{"key":"e_1_3_2_2_24_1","volume-title":"Pytorch: An imperative style, high-performance deep learning library. NeurIPS","author":"Paszke Adam","year":"2019","unstructured":"Adam Paszke , Sam Gross , Francisco Massa , Adam Lerer , James Bradbury , Gregory Chanan , Trevor Killeen , Zeming Lin , Natalia Gimelshein , Luca Antiga , 2019 . Pytorch: An imperative style, high-performance deep learning library. NeurIPS (2019). Adam Paszke, Sam Gross, Francisco Massa, Adam Lerer, James Bradbury, Gregory Chanan, Trevor Killeen, Zeming Lin, Natalia Gimelshein, Luca Antiga, et al. 2019. Pytorch: An imperative style, high-performance deep learning library. NeurIPS (2019)."},{"key":"e_1_3_2_2_25_1","volume-title":"Meld: A multimodal multi-party dataset for emotion recognition in conversations. arXiv preprint arXiv:1810.02508","author":"Poria Soujanya","year":"2018","unstructured":"Soujanya Poria , Devamanyu Hazarika , Navonil Majumder , Gautam Naik , Erik Cambria , and Rada Mihalcea . 2018 . Meld: A multimodal multi-party dataset for emotion recognition in conversations. arXiv preprint arXiv:1810.02508 (2018). Soujanya Poria, Devamanyu Hazarika, Navonil Majumder, Gautam Naik, Erik Cambria, and Rada Mihalcea. 2018. Meld: A multimodal multi-party dataset for emotion recognition in conversations. arXiv preprint arXiv:1810.02508 (2018)."},{"key":"e_1_3_2_2_26_1","volume-title":"Learning multi-level deep representations for image emotion classification. Neural Processing Letters","author":"Rao Tianrong","year":"2020","unstructured":"Tianrong Rao , Xiaoxu Li , and Min Xu. 2020. Learning multi-level deep representations for image emotion classification. Neural Processing Letters ( 2020 ). Tianrong Rao, Xiaoxu Li, and Min Xu. 2020. Learning multi-level deep representations for image emotion classification. Neural Processing Letters (2020)."},{"key":"e_1_3_2_2_27_1","volume-title":"Very deep convolutional networks for large-scale image recognition. arXiv preprint arXiv:1409.1556","author":"Simonyan Karen","year":"2014","unstructured":"Karen Simonyan and Andrew Zisserman . 2014. Very deep convolutional networks for large-scale image recognition. arXiv preprint arXiv:1409.1556 ( 2014 ). Karen Simonyan and Andrew Zisserman. 2014. Very deep convolutional networks for large-scale image recognition. arXiv preprint arXiv:1409.1556 (2014)."},{"key":"e_1_3_2_2_28_1","first-page":"27","article-title":"Emoticon, emoji, and sticker use in computer-mediated communication: A review of theories and research findings","volume":"13","author":"Tang Ying","year":"2019","unstructured":"Ying Tang and Khe Foon Hew . 2019 . Emoticon, emoji, and sticker use in computer-mediated communication: A review of theories and research findings . International Journal of Communication , Vol. 13 (2019), 27 . Ying Tang and Khe Foon Hew. 2019. Emoticon, emoji, and sticker use in computer-mediated communication: A review of theories and research findings. International Journal of Communication, Vol. 13 (2019), 27.","journal-title":"International Journal of Communication"},{"key":"e_1_3_2_2_29_1","volume-title":"Vistanet: Visual aspect attention network for multimodal sentiment analysis. In AAAI.","author":"Truong Quoc-Tuan","year":"2019","unstructured":"Quoc-Tuan Truong and Hady W Lauw . 2019 . Vistanet: Visual aspect attention network for multimodal sentiment analysis. In AAAI. Quoc-Tuan Truong and Hady W Lauw. 2019. Vistanet: Visual aspect attention network for multimodal sentiment analysis. In AAAI."},{"key":"e_1_3_2_2_30_1","volume-title":"Attention is all you need. NeurIPS","author":"Vaswani Ashish","year":"2017","unstructured":"Ashish Vaswani , Noam Shazeer , Niki Parmar , Jakob Uszkoreit , Llion Jones , Aidan N Gomez , \u0141ukasz Kaiser , and Illia Polosukhin . 2017. Attention is all you need. NeurIPS ( 2017 ). Ashish Vaswani, Noam Shazeer, Niki Parmar, Jakob Uszkoreit, Llion Jones, Aidan N Gomez, \u0141ukasz Kaiser, and Illia Polosukhin. 2017. Attention is all you need. NeurIPS (2017)."},{"key":"e_1_3_2_2_31_1","doi-asserted-by":"crossref","unstructured":"Wenhai Wang Enze Xie Xiang Li Deng-Ping Fan Kaitao Song Ding Liang Tong Lu Ping Luo and Ling Shao. 2021. Pyramid vision transformer: A versatile backbone for dense prediction without convolutions. In ICCV.  Wenhai Wang Enze Xie Xiang Li Deng-Ping Fan Kaitao Song Ding Liang Tong Lu Ping Luo and Ling Shao. 2021. Pyramid vision transformer: A versatile backbone for dense prediction without convolutions. In ICCV.","DOI":"10.1109\/ICCV48922.2021.00061"},{"key":"e_1_3_2_2_32_1","unstructured":"Nan Xu Wenji Mao and Guandan Chen. [n.d.]. Multi-interactive memory network for aspect based multimodal sentiment analysis. In AAAI.  Nan Xu Wenji Mao and Guandan Chen. [n.d.]. Multi-interactive memory network for aspect based multimodal sentiment analysis. In AAAI."},{"key":"e_1_3_2_2_33_1","doi-asserted-by":"publisher","DOI":"10.1109\/TIP.2021.3118983"},{"key":"e_1_3_2_2_34_1","doi-asserted-by":"crossref","unstructured":"Jingyuan Yang Jie Li Leida Li Xiumei Wang and Xinbo Gao. 2021b. A circular-structured representation for visual emotion distribution learning. In CVPR.  Jingyuan Yang Jie Li Leida Li Xiumei Wang and Xinbo Gao. 2021b. A circular-structured representation for visual emotion distribution learning. In CVPR.","DOI":"10.1109\/CVPR46437.2021.00422"},{"key":"e_1_3_2_2_35_1","doi-asserted-by":"publisher","DOI":"10.1109\/TIP.2021.3106813"},{"key":"e_1_3_2_2_36_1","doi-asserted-by":"crossref","unstructured":"Jufeng Yang Dongyu She Yu-Kun Lai Paul L Rosin and Ming-Hsuan Yang. 2018b. Weakly supervised coupled networks for visual sentiment analysis. In CVPR.  Jufeng Yang Dongyu She Yu-Kun Lai Paul L Rosin and Ming-Hsuan Yang. 2018b. Weakly supervised coupled networks for visual sentiment analysis. In CVPR.","DOI":"10.1109\/CVPR.2018.00791"},{"key":"e_1_3_2_2_37_1","doi-asserted-by":"crossref","unstructured":"Jufeng Yang Dongyu She Yu-Kun Lai and Ming-Hsuan Yang. 2018a. Retrieving and classifying affective images via deep metric learning. In AAAI.  Jufeng Yang Dongyu She Yu-Kun Lai and Ming-Hsuan Yang. 2018a. Retrieving and classifying affective images via deep metric learning. In AAAI.","DOI":"10.1609\/aaai.v32i1.11275"},{"key":"e_1_3_2_2_38_1","doi-asserted-by":"crossref","unstructured":"Jufeng Yang Dongyu She and Ming Sun. 2017. Joint image emotion classification and distribution learning via deep convolutional neural network. In IJCAI.  Jufeng Yang Dongyu She and Ming Sun. 2017. Joint image emotion classification and distribution learning via deep convolutional neural network. In IJCAI.","DOI":"10.24963\/ijcai.2017\/456"},{"key":"e_1_3_2_2_39_1","doi-asserted-by":"publisher","DOI":"10.1109\/TMM.2018.2803520"},{"key":"e_1_3_2_2_40_1","doi-asserted-by":"publisher","DOI":"10.1109\/TMM.2020.3035277"},{"key":"e_1_3_2_2_41_1","doi-asserted-by":"crossref","unstructured":"Xingxu Yao Dongyu She Sicheng Zhao Jie Liang Yu-Kun Lai and Jufeng Yang. 2019. Attention-aware polarity sensitive embedding for affective image retrieval. In ICCV.  Xingxu Yao Dongyu She Sicheng Zhao Jie Liang Yu-Kun Lai and Jufeng Yang. 2019. Attention-aware polarity sensitive embedding for affective image retrieval. In ICCV.","DOI":"10.1109\/ICCV.2019.00123"},{"key":"e_1_3_2_2_42_1","unstructured":"Quanzeng You Jiebo Luo Hailin Jin and Jianchao Yang. 2015. Robust image sentiment analysis using progressively trained and domain transferred deep networks. In AAAI.  Quanzeng You Jiebo Luo Hailin Jin and Jianchao Yang. 2015. Robust image sentiment analysis using progressively trained and domain transferred deep networks. In AAAI."},{"key":"e_1_3_2_2_43_1","unstructured":"Quanzeng You Jiebo Luo Hailin Jin and Jianchao Yang. 2016. Building a large scale dataset for image emotion recognition: The fine print and the benchmark. In AAAI.  Quanzeng You Jiebo Luo Hailin Jin and Jianchao Yang. 2016. Building a large scale dataset for image emotion recognition: The fine print and the benchmark. In AAAI."},{"key":"e_1_3_2_2_44_1","volume-title":"Tensor fusion network for multimodal sentiment analysis. arXiv preprint arXiv:1707.07250","author":"Zadeh Amir","year":"2017","unstructured":"Amir Zadeh , Minghai Chen , Soujanya Poria , Erik Cambria , and Louis-Philippe Morency . 2017. Tensor fusion network for multimodal sentiment analysis. arXiv preprint arXiv:1707.07250 ( 2017 ). Amir Zadeh, Minghai Chen, Soujanya Poria, Erik Cambria, and Louis-Philippe Morency. 2017. Tensor fusion network for multimodal sentiment analysis. arXiv preprint arXiv:1707.07250 (2017)."},{"key":"e_1_3_2_2_45_1","unstructured":"Amir Zadeh and Paul Pu. 2018. Multimodal language analysis in the wild: Cmu-mosei dataset and interpretable dynamic fusion graph. In ACL.  Amir Zadeh and Paul Pu. 2018. Multimodal language analysis in the wild: Cmu-mosei dataset and interpretable dynamic fusion graph. In ACL."},{"key":"e_1_3_2_2_46_1","doi-asserted-by":"publisher","DOI":"10.1109\/TMM.2019.2928998"},{"key":"e_1_3_2_2_47_1","doi-asserted-by":"crossref","unstructured":"Sicheng Zhao Yue Gao Xiaolei Jiang Hongxun Yao Tat-Seng Chua and Xiaoshuai Sun. 2014. Exploring principles-of-art features for image emotion recognition. In ACM MM.  Sicheng Zhao Yue Gao Xiaolei Jiang Hongxun Yao Tat-Seng Chua and Xiaoshuai Sun. 2014. Exploring principles-of-art features for image emotion recognition. In ACM MM.","DOI":"10.1145\/2647868.2654930"},{"key":"e_1_3_2_2_48_1","doi-asserted-by":"publisher","DOI":"10.1145\/3343031.3351062"},{"key":"e_1_3_2_2_49_1","doi-asserted-by":"crossref","unstructured":"Sicheng Zhao Hongxun Yao Yue Gao Rongrong Ji Wenlong Xie Xiaolei Jiang and Tat-Seng Chua. 2016. Predicting personalized emotion perceptions of social images. In ACM MM.  Sicheng Zhao Hongxun Yao Yue Gao Rongrong Ji Wenlong Xie Xiaolei Jiang and Tat-Seng Chua. 2016. Predicting personalized emotion perceptions of social images. In ACM MM.","DOI":"10.1145\/2964284.2964289"},{"key":"e_1_3_2_2_50_1","volume-title":"Affective image content analysis: Two decades review and new perspectives","author":"Zhao Sicheng","year":"2021","unstructured":"Sicheng Zhao , Xingxu Yao , Jufeng Yang , Guoli Jia , Guiguang Ding , Tat-Seng Chua , Bjoern W Schuller , and Kurt Keutzer . 2021. Affective image content analysis: Two decades review and new perspectives . IEEE Transactions on Pattern Analysis and Machine Intelligence ( 2021 ). Sicheng Zhao, Xingxu Yao, Jufeng Yang, Guoli Jia, Guiguang Ding, Tat-Seng Chua, Bjoern W Schuller, and Kurt Keutzer. 2021. Affective image content analysis: Two decades review and new perspectives. IEEE Transactions on Pattern Analysis and Machine Intelligence (2021)."},{"key":"e_1_3_2_2_51_1","doi-asserted-by":"publisher","DOI":"10.1109\/TPAMI.2017.2723009"},{"key":"e_1_3_2_2_52_1","doi-asserted-by":"crossref","unstructured":"Xinge Zhu Liang Li Weigang Zhang Tianrong Rao Min Xu Qingming Huang and Dong Xu. 2017. Dependency exploitation: A unified CNN-RNN approach for visual emotion recognition. In IJCAI.  Xinge Zhu Liang Li Weigang Zhang Tianrong Rao Min Xu Qingming Huang and Dong Xu. 2017. Dependency exploitation: A unified CNN-RNN approach for visual emotion recognition. In IJCAI.","DOI":"10.24963\/ijcai.2017\/503"}],"event":{"name":"MM '22: The 30th ACM International Conference on Multimedia","location":"Lisboa Portugal","acronym":"MM '22","sponsor":["SIGMM ACM Special Interest Group on Multimedia"]},"container-title":["Proceedings of the 30th ACM International Conference on Multimedia"],"original-title":[],"link":[{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/3503161.3548407","content-type":"unspecified","content-version":"vor","intended-application":"text-mining"},{"URL":"https:\/\/dl.acm.org\/doi\/pdf\/10.1145\/3503161.3548407","content-type":"unspecified","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2025,6,17]],"date-time":"2025-06-17T17:49:17Z","timestamp":1750182557000},"score":1,"resource":{"primary":{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/3503161.3548407"}},"subtitle":["A Large-Scale Dataset for Sticker Emotion Recognition"],"short-title":[],"issued":{"date-parts":[[2022,10,10]]},"references-count":52,"alternative-id":["10.1145\/3503161.3548407","10.1145\/3503161"],"URL":"https:\/\/doi.org\/10.1145\/3503161.3548407","relation":{},"subject":[],"published":{"date-parts":[[2022,10,10]]},"assertion":[{"value":"2022-10-10","order":2,"name":"published","label":"Published","group":{"name":"publication_history","label":"Publication History"}}]}}