{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2026,1,21]],"date-time":"2026-01-21T17:39:39Z","timestamp":1769017179677,"version":"3.49.0"},"publisher-location":"New York, NY, USA","reference-count":32,"publisher":"ACM","license":[{"start":{"date-parts":[[2019,10,15]],"date-time":"2019-10-15T00:00:00Z","timestamp":1571097600000},"content-version":"vor","delay-in-days":0,"URL":"https:\/\/www.acm.org\/publications\/policies\/copyright_policy#Background"}],"funder":[{"name":"National Institutes of Health","award":["R01LM011834"],"award-info":[{"award-number":["R01LM011834"]}]}],"content-domain":{"domain":["dl.acm.org"],"crossmark-restriction":true},"short-container-title":[],"published-print":{"date-parts":[[2019,10,15]]},"DOI":"10.1145\/3343031.3351039","type":"proceedings-article","created":{"date-parts":[[2019,10,21]],"date-time":"2019-10-21T16:32:26Z","timestamp":1571675546000},"page":"157-166","update-policy":"https:\/\/doi.org\/10.1145\/crossmark-policy","source":"Crossref","is-referenced-by-count":16,"title":["Mutual Correlation Attentive Factors in Dyadic Fusion Networks for Speech Emotion Recognition"],"prefix":"10.1145","author":[{"given":"Yue","family":"Gu","sequence":"first","affiliation":[{"name":"Rutgers University, Piscataway, NJ, USA"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Xinyu","family":"Lyu","sequence":"additional","affiliation":[{"name":"Rutgers University, Piscataway, NJ, USA"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Weijia","family":"Sun","sequence":"additional","affiliation":[{"name":"Rutgers University, Piscataway, NJ, USA"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Weitian","family":"Li","sequence":"additional","affiliation":[{"name":"Rutgers University, Piscataway, NJ, USA"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Shuhong","family":"Chen","sequence":"additional","affiliation":[{"name":"Rutgers University, Piscataway, NJ, USA"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Xinyu","family":"Li","sequence":"additional","affiliation":[{"name":"Rutgers University, Piscataway, NJ, USA"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Ivan","family":"Marsic","sequence":"additional","affiliation":[{"name":"Rutgers University, Piscataway, NJ, USA"}],"role":[{"role":"author","vocabulary":"crossref"}]}],"member":"320","published-online":{"date-parts":[[2019,10,15]]},"reference":[{"key":"e_1_3_2_1_1_1","doi-asserted-by":"crossref","unstructured":"Amir Zadeh Minghai Chen Soujanya Poria Erik Cambria and Louis-Philippe Morency. 2017. \"Tensor fusion network for multimodal sentiment analysis.\" arXiv preprint arXiv:1707.07250 (2017).  Amir Zadeh Minghai Chen Soujanya Poria Erik Cambria and Louis-Philippe Morency. 2017. \"Tensor fusion network for multimodal sentiment analysis.\" arXiv preprint arXiv:1707.07250 (2017).","DOI":"10.18653\/v1\/D17-1115"},{"key":"e_1_3_2_1_2_1","doi-asserted-by":"crossref","unstructured":"Soujanya Poria Erik Cambria and Alexander Gelbukh. 2015. \"Deep convolutional neural network textual features and multiple kernel learning for utter-ance-level multimodal sentiment analysis.\" In Proceedings of the 2015 conference on empirical methods in natural language processing pp. (2015). 2539--2544.  Soujanya Poria Erik Cambria and Alexander Gelbukh. 2015. \"Deep convolutional neural network textual features and multiple kernel learning for utter-ance-level multimodal sentiment analysis.\" In Proceedings of the 2015 conference on empirical methods in natural language processing pp. (2015). 2539--2544.","DOI":"10.18653\/v1\/D15-1303"},{"key":"e_1_3_2_1_3_1","first-page":"2225","volume-title":"Long Papers) (Vol. 1","author":"Gu Yue","unstructured":"Yue Gu , Kangning Yang , Shiyu Fu , Shuhong Chen , Xinyu Li , and Ivan Mar-sic. 2018. \"Multimodal affective analysis using hierarchical attention strategy with word-level alignment.\" In Proceedings of the 56th Annual Meeting of the Association for Computational Linguistics (Volume 1 : Long Papers) (Vol. 1 , pp. 2225 -- 2235 ). Yue Gu, Kangning Yang, Shiyu Fu, Shuhong Chen, Xinyu Li, and Ivan Mar-sic. 2018. \"Multimodal affective analysis using hierarchical attention strategy with word-level alignment.\" In Proceedings of the 56th Annual Meeting of the Association for Computational Linguistics (Volume 1: Long Papers) (Vol. 1, pp. 2225--2235)."},{"key":"e_1_3_2_1_4_1","first-page":"873","volume-title":"Long Papers)","author":"Poria Soujanya","unstructured":"Soujanya Poria , Erik Cambria , Devamanyu Hazarika , Navonil Majumder , Amir Zadeh , and Louis-Philippe Morency. 2017. \"Context-dependent sentiment analysis in user-generated videos.\" In Proceedings of the 55th Annual Meeting of the Association for Computational Linguistics (Volume 1 : Long Papers) , pp. 873 -- 883 . Soujanya Poria, Erik Cambria, Devamanyu Hazarika, Navonil Majumder, Amir Zadeh, and Louis-Philippe Morency. 2017. \"Context-dependent sentiment analysis in user-generated videos.\" In Proceedings of the 55th Annual Meeting of the Association for Computational Linguistics (Volume 1: Long Papers), pp. 873--883."},{"key":"e_1_3_2_1_5_1","volume-title":"Soujanya Poria, Prateek Vij, Erik Cambria, and Louis-Philippe Morency.","author":"Zadeh Amir","year":"2018","unstructured":"Amir Zadeh , Paul Pu Liang , Soujanya Poria, Prateek Vij, Erik Cambria, and Louis-Philippe Morency. 2018 . \"Multi-attention recurrent network for human communication comprehension.\" In Thirty-Second AAAI Conference on Artifi-cial Intelligence . (2018). Amir Zadeh, Paul Pu Liang, Soujanya Poria, Prateek Vij, Erik Cambria, and Louis-Philippe Morency. 2018. \"Multi-attention recurrent network for human communication comprehension.\" In Thirty-Second AAAI Conference on Artifi-cial Intelligence. (2018)."},{"key":"e_1_3_2_1_6_1","volume-title":"ACM","author":"Gu Yue","year":"2018","unstructured":"Yue Gu , Xinyu Li , Kaixiang Huang , Shiyu Fu , Kangning Yang , Shuhong Chen , Moliang Zhou , and Ivan Marsic . 2018 . \" Human conversation analysis using attentive multimodal networks with hierarchical encoder-decoder.\" In 2018 ACM Multimedia Conference on Multimedia Conference, pp. 537--545 . ACM , (2018). Yue Gu, Xinyu Li, Kaixiang Huang, Shiyu Fu, Kangning Yang, Shuhong Chen, Moliang Zhou, and Ivan Marsic. 2018. \"Human conversation analysis using attentive multimodal networks with hierarchical encoder-decoder.\" In 2018 ACM Multimedia Conference on Multimedia Conference, pp. 537--545. ACM, (2018)."},{"key":"e_1_3_2_1_7_1","volume-title":"IEEE","author":"Poria Soujanya","year":"2016","unstructured":"Soujanya Poria , Iti Chaturvedi , Erik Cambria , and Amir Hussain . 2016 . \" Convolutional MKL based multimodal emotion recognition and sentiment analysis.\" In 2016 IEEE 16th international conference on data mining (ICDM), pp. 439--448 . IEEE , (2016). Soujanya Poria, Iti Chaturvedi, Erik Cambria, and Amir Hussain. 2016. \"Convolutional MKL based multimodal emotion recognition and sentiment analysis.\" In 2016 IEEE 16th international conference on data mining (ICDM), pp. 439--448. IEEE, (2016)."},{"key":"e_1_3_2_1_8_1","volume-title":"ACM","author":"Chen Minghai","year":"2017","unstructured":"Minghai Chen , Sen Wang , Paul Pu Liang , Tadas Baltruaitis , Amir Zadeh , and Louis-Philippe Morency . 2017 . \" Multimodal sentiment analysis with word-level fusion and reinforcement learning.\" In Proceedings of the 19th ACM International Conference on Multimodal Interaction, pp. 163--171 . ACM , (2017). Minghai Chen, Sen Wang, Paul Pu Liang, Tadas Baltruaitis, Amir Zadeh, and Louis-Philippe Morency. 2017. \"Multimodal sentiment analysis with word-level fusion and reinforcement learning.\" In Proceedings of the 19th ACM International Conference on Multimodal Interaction, pp. 163--171. ACM, (2017)."},{"key":"e_1_3_2_1_9_1","volume-title":"Narayanan","author":"Busso Carlos","year":"2008","unstructured":"Carlos Busso , Murtaza Bulut , Chi-Chun Lee , Abe Kazemzadeh , Emily Mower , Samuel Kim , Jeannette N. Chang , Sungbok Lee , and Shrikanth S . Narayanan . 2008 . \"IEMOCAP: Interactive emotional dyadic motion capture database.\" Language resources and evaluation 42, no. 4 (2008): 335. Carlos Busso, Murtaza Bulut, Chi-Chun Lee, Abe Kazemzadeh, Emily Mower, Samuel Kim, Jeannette N. Chang, Sungbok Lee, and Shrikanth S. Narayanan. 2008. \"IEMOCAP: Interactive emotional dyadic motion capture database.\" Language resources and evaluation 42, no. 4 (2008): 335."},{"key":"e_1_3_2_1_10_1","first-page":"5998","volume-title":"ukasz Kaiser, and Illia Polosukhin","author":"Vaswani Ashish","year":"2017","unstructured":"Ashish Vaswani , Noam Shazeer , Niki Parmar , Jakob Uszkoreit , Llion Jones , Aidan N. Gomez , ukasz Kaiser, and Illia Polosukhin . 2017 . \"Attention is all you need.\" In Advances in neural information processing systems, pp. 5998 -- 6008 . (2017). Ashish Vaswani, Noam Shazeer, Niki Parmar, Jakob Uszkoreit, Llion Jones, Aidan N. Gomez, ukasz Kaiser, and Illia Polosukhin. 2017. \"Attention is all you need.\" In Advances in neural information processing systems, pp. 5998--6008. (2017)."},{"key":"e_1_3_2_1_11_1","volume-title":"IEEE","author":"Poria Soujanya","year":"2017","unstructured":"Soujanya Poria , Erik Cambria , Devamanyu Hazarika , Navonil Mazumder , Amir Zadeh , and Louis-Philippe Morency . 2017 . \" Multi-level multiple attentions for contextual multimodal sentiment analysis.\" In 2017 IEEE International Conference on Data Mining (ICDM), pp. 1033--1038 . IEEE , (2017). Soujanya Poria, Erik Cambria, Devamanyu Hazarika, Navonil Mazumder, Amir Zadeh, and Louis-Philippe Morency. 2017. \"Multi-level multiple attentions for contextual multimodal sentiment analysis.\" In 2017 IEEE International Conference on Data Mining (ICDM), pp. 1033--1038. IEEE, (2017)."},{"key":"e_1_3_2_1_12_1","volume-title":"A multimodal multi-party dataset for emotion recognition in conversations.\" arXiv preprint arXiv:1810.02508","author":"Poria Soujanya","year":"2018","unstructured":"Soujanya Poria , Devamanyu Hazarika , Navonil Majumder , Gautam Naik , Erik Cambria , and Rada Mihalcea . 2018. \"Meld : A multimodal multi-party dataset for emotion recognition in conversations.\" arXiv preprint arXiv:1810.02508 ( 2018 ). Soujanya Poria, Devamanyu Hazarika, Navonil Majumder, Gautam Naik, Erik Cambria, and Rada Mihalcea. 2018. \"Meld: A multimodal multi-party dataset for emotion recognition in conversations.\" arXiv preprint arXiv:1810.02508 (2018)."},{"key":"e_1_3_2_1_13_1","first-page":"601","volume-title":"12th Australasian Int. Conf. on Speech Science and Technology SST 2008","author":"Seppi Dino","year":"2008","unstructured":"Dino Seppi , Anton Batliner , Bj\u00f6rn Schuller , Stefan Steidl , Thurid Vogt , Johannes Wagner , Laurence Devillers , Laurence Vidrascu , Noam Amir , and Vered Aharonson . 2008 . \" Patterns, prototypes, performance: classifying emotional user states.\" In Proc. 9th Interspeech 2008 incorp . 12th Australasian Int. Conf. on Speech Science and Technology SST 2008 , Brisbane, Australia , pp. 601 -- 604 . (2008). Dino Seppi, Anton Batliner, Bj\u00f6rn Schuller, Stefan Steidl, Thurid Vogt, Johannes Wagner, Laurence Devillers, Laurence Vidrascu, Noam Amir, and Vered Aharonson. 2008. \"Patterns, prototypes, performance: classifying emotional user states.\" In Proc. 9th Interspeech 2008 incorp. 12th Australasian Int. Conf. on Speech Science and Technology SST 2008, Brisbane, Australia, pp. 601--604. (2008)."},{"key":"e_1_3_2_1_14_1","first-page":"485","volume-title":"audio and lexical indicators of affect in spontaneous conversation via particle filtering.\" In Proceedings of the 14th ACM international conference on Multimodal interaction","author":"Savran Arman","year":"2012","unstructured":"Arman Savran , Houwei Cao , Miraj Shah , Ani Nenkova , and Ragini Verma . 2012. \" Combining video , audio and lexical indicators of affect in spontaneous conversation via particle filtering.\" In Proceedings of the 14th ACM international conference on Multimodal interaction , pp. 485 -- 492 . ACM , ( 2012 ). Arman Savran, Houwei Cao, Miraj Shah, Ani Nenkova, and Ragini Verma. 2012. \"Combining video, audio and lexical indicators of affect in spontaneous conversation via particle filtering.\" In Proceedings of the 14th ACM international conference on Multimodal interaction, pp. 485--492. ACM, (2012)."},{"key":"e_1_3_2_1_15_1","volume-title":"ACM","author":"Eyben Florian","year":"2010","unstructured":"Florian Eyben , Martin W\u00f6llmer , and Bj\u00f6rn Schuller . 2010 . \" Opensmile: the munich versatile and fast open-source audio feature extractor.\" In Proceedings of the 18th ACM international conference on Multimedia, pp. 1459--1462 . ACM , (2010). Florian Eyben, Martin W\u00f6llmer, and Bj\u00f6rn Schuller. 2010. \"Opensmile: the munich versatile and fast open-source audio feature extractor.\" In Proceedings of the 18th ACM international conference on Multimedia, pp. 1459--1462. ACM, (2010)."},{"key":"e_1_3_2_1_16_1","first-page":"960","volume-title":"speech and signal processing (icassp)","author":"Degottex Gilles","year":"2014","unstructured":"Gilles Degottex , John Kane , Thomas Drugman , Tuomo Raitio , and Stefan Scherer . 2014. \" COVAREP-A collaborative voice analysis repository for speech technologies.\" In 2014 ieee international conference on acoustics , speech and signal processing (icassp) , pp. 960 -- 964 . IEEE , ( 2014 ). Gilles Degottex, John Kane, Thomas Drugman, Tuomo Raitio, and Stefan Scherer. 2014. \"COVAREP-A collaborative voice analysis repository for speech technologies.\" In 2014 ieee international conference on acoustics, speech and signal processing (icassp), pp. 960--964. IEEE, (2014)."},{"key":"e_1_3_2_1_17_1","first-page":"5079","volume-title":"Speech and Signal Processing (ICASSP)","author":"Gu Yue","year":"2018","unstructured":"Yue Gu , Shuhong Chen , and Ivan Marsic . 2018 . \" Deep Multimodal learning for emotion recognition in spoken language.\" In 2018 IEEE International Conference on Acoustics , Speech and Signal Processing (ICASSP) , pp. 5079 -- 5083 . IEEE, (2018). Yue Gu, Shuhong Chen, and Ivan Marsic. 2018. \"Deep Multimodal learning for emotion recognition in spoken language.\" In 2018 IEEE International Conference on Acoustics, Speech and Signal Processing (ICASSP), pp. 5079--5083. IEEE, (2018)."},{"key":"e_1_3_2_1_18_1","volume-title":"Springer","author":"Rajagopalan Shyam Sundar","year":"2016","unstructured":"Shyam Sundar Rajagopalan , Louis-Philippe Morency , Tadas Baltrusaitis , and Roland Goecke . 2016 . \" Extending long short-term memory for multi-view structured learning.\" In European Conference on Computer Vision, pp. 338--353 . Springer , Cham , (2016). Shyam Sundar Rajagopalan, Louis-Philippe Morency, Tadas Baltrusaitis, and Roland Goecke. 2016. \"Extending long short-term memory for multi-view structured learning.\" In European Conference on Computer Vision, pp. 338--353. Springer, Cham, (2016)."},{"key":"e_1_3_2_1_19_1","doi-asserted-by":"crossref","unstructured":"Martin W\u00f6llmer Felix Weninger Tobias Knaup Bj\u00f6rn Schuller Congkai Sun Kenji Sagae and Louis-Philippe Morency. 2013. \"Youtube movie reviews: Sentiment analysis in an audio-visual context.\" IEEE Intelligent Systems 28 no. 3 (2013): 46--53.  Martin W\u00f6llmer Felix Weninger Tobias Knaup Bj\u00f6rn Schuller Congkai Sun Kenji Sagae and Louis-Philippe Morency. 2013. \"Youtube movie reviews: Sentiment analysis in an audio-visual context.\" IEEE Intelligent Systems 28 no. 3 (2013): 46--53.","DOI":"10.1109\/MIS.2013.34"},{"key":"e_1_3_2_1_20_1","volume-title":"Paul Pu Liang, Amir Zadeh, and Louis-Philippe Morency.","author":"Liu Zhun","year":"2018","unstructured":"Zhun Liu , Ying Shen , Varun Bharadhwaj Lakshminarasimhan , Paul Pu Liang, Amir Zadeh, and Louis-Philippe Morency. 2018 . \"Efficient low-rank multi-modal fusion with modality-specific factors.\" arXiv preprint arXiv:1806.00064 (2018). Zhun Liu, Ying Shen, Varun Bharadhwaj Lakshminarasimhan, Paul Pu Liang, Amir Zadeh, and Louis-Philippe Morency. 2018. \"Efficient low-rank multi-modal fusion with modality-specific factors.\" arXiv preprint arXiv:1806.00064 (2018)."},{"key":"e_1_3_2_1_21_1","volume-title":"ACM","author":"Liang Paul Pu","year":"2018","unstructured":"Paul Pu Liang , Amir Zadeh , and Louis-Philippe Morency . 2018 . \" Multimodal local-global ranking fusion for emotion recognition.\" In Proceedings of the 2018 on International Conference on Multimodal Interaction, pp. 472--476 . ACM , (2018). Paul Pu Liang, Amir Zadeh, and Louis-Philippe Morency. 2018. \"Multimodal local-global ranking fusion for emotion recognition.\" In Proceedings of the 2018 on International Conference on Multimodal Interaction, pp. 472--476. ACM, (2018)."},{"key":"e_1_3_2_1_22_1","first-page":"1532","volume-title":"Global vectors for word representation.\" In Proceedings of the 2014 conference on empirical methods in natural language processing (EMNLP)","author":"Pennington Jeffrey","year":"2014","unstructured":"Jeffrey Pennington , Richard Socher , and Christopher Manning . 2014. \"Glove : Global vectors for word representation.\" In Proceedings of the 2014 conference on empirical methods in natural language processing (EMNLP) , pp. 1532 -- 1543 . ( 2014 ). Jeffrey Pennington, Richard Socher, and Christopher Manning. 2014. \"Glove: Global vectors for word representation.\" In Proceedings of the 2014 conference on empirical methods in natural language processing (EMNLP), pp. 1532--1543. (2014)."},{"key":"e_1_3_2_1_23_1","volume-title":"Accelerating deep network training by reducing internal covariate shift.\" arXiv preprint arXiv:1502.03167","author":"Ioffe Sergey","year":"2015","unstructured":"Sergey Ioffe , and Christian Szegedy . 2015. \" Batch normalization : Accelerating deep network training by reducing internal covariate shift.\" arXiv preprint arXiv:1502.03167 ( 2015 ). Sergey Ioffe, and Christian Szegedy. 2015. \"Batch normalization: Accelerating deep network training by reducing internal covariate shift.\" arXiv preprint arXiv:1502.03167 (2015)."},{"key":"e_1_3_2_1_24_1","unstructured":"Yuji Tokozume Yoshitaka Ushiku and Tatsuya Harada. 2017. \"Learning from between-class examples for deep sound recognition.\" arXiv preprint arXiv:1711.10282 (2017).  Yuji Tokozume Yoshitaka Ushiku and Tatsuya Harada. 2017. \"Learning from between-class examples for deep sound recognition.\" arXiv preprint arXiv:1711.10282 (2017)."},{"key":"e_1_3_2_1_25_1","volume-title":"An emotion corpus of multi-party conversations.\" arXiv pre-print arXiv:1802.08379","author":"Chen Sheng-Yeh","year":"2018","unstructured":"Sheng-Yeh Chen , Chao-Chun Hsu , Chuan-Chun Kuo , and Lun-Wei Ku. 2018. \" Emotionlines : An emotion corpus of multi-party conversations.\" arXiv pre-print arXiv:1802.08379 ( 2018 ). Sheng-Yeh Chen, Chao-Chun Hsu, Chuan-Chun Kuo, and Lun-Wei Ku. 2018. \"Emotionlines: An emotion corpus of multi-party conversations.\" arXiv pre-print arXiv:1802.08379 (2018)."},{"key":"e_1_3_2_1_26_1","doi-asserted-by":"crossref","unstructured":"Ver\u00f3nica P\u00e9rez Rosas Rada Mihalcea and Louis-Philippe Morency. 2013. \"Multimodal sentiment analysis of spanish online videos.\" IEEE Intelligent Systems 28 no. 3 (2013): 38--45.  Ver\u00f3nica P\u00e9rez Rosas Rada Mihalcea and Louis-Philippe Morency. 2013. \"Multimodal sentiment analysis of spanish online videos.\" IEEE Intelligent Systems 28 no. 3 (2013): 38--45.","DOI":"10.1109\/MIS.2013.9"},{"key":"e_1_3_2_1_27_1","doi-asserted-by":"crossref","unstructured":"Leo Breiman. 2001. \"Random forests.\" Machine learning 45 no. 1 (2001): 5--32.  Leo Breiman. 2001. \"Random forests.\" Machine learning 45 no. 1 (2001): 5--32.","DOI":"10.1023\/A:1010933404324"},{"key":"e_1_3_2_1_28_1","unstructured":"Fran\u00e7ois Chollet. 2015. \"Keras.\" (2015).  Fran\u00e7ois Chollet. 2015. \"Keras.\" (2015)."},{"key":"e_1_3_2_1_29_1","first-page":"265","volume-title":"Matthieu Devin et al","author":"Abadi Mart\u00edn","year":"2016","unstructured":"Mart\u00edn Abadi , Paul Barham , Jianmin Chen , Zhifeng Chen , Andy Davis , Jeffrey Dean , Matthieu Devin et al . 2016 . \"Tensorflow: A system for large-scale ma-chine learning.\" In 12th {USENIX} Symposium on Operating Systems Design and Implementation ( {OSDI} 16), pp. 265 -- 283 . (2016). Mart\u00edn Abadi, Paul Barham, Jianmin Chen, Zhifeng Chen, Andy Davis, Jeffrey Dean, Matthieu Devin et al. 2016. \"Tensorflow: A system for large-scale ma-chine learning.\" In 12th {USENIX} Symposium on Operating Systems Design and Implementation ({OSDI} 16), pp. 265--283. (2016)."},{"key":"e_1_3_2_1_30_1","volume-title":"A method for stochastic optimization.\" arXiv preprint arXiv:1412.6980","author":"Kingma Diederik P.","year":"2014","unstructured":"Diederik P. Kingma , and Jimmy Ba. 2014. \" Adam : A method for stochastic optimization.\" arXiv preprint arXiv:1412.6980 ( 2014 ). Diederik P. Kingma, and Jimmy Ba. 2014. \"Adam: A method for stochastic optimization.\" arXiv preprint arXiv:1412.6980 (2014)."},{"key":"e_1_3_2_1_31_1","first-page":"2379","article-title":"Hybrid Attention based Multimodal Network for Spoken Language Classification","author":"Gu Yue","year":"2018","unstructured":"Yue Gu , Kangning Yang , Shiyu Fu , Shuhong Chen , Xinyu Li , and Ivan Mar-sic. 2018 . \" Hybrid Attention based Multimodal Network for Spoken Language Classification .\" In Proceedings of the 27th International Conference on Com-putational Linguistics , pp. 2379 -- 2390 . (2018). Yue Gu, Kangning Yang, Shiyu Fu, Shuhong Chen, Xinyu Li, and Ivan Mar-sic. 2018. \"Hybrid Attention based Multimodal Network for Spoken Language Classification.\" In Proceedings of the 27th International Conference on Com-putational Linguistics, pp. 2379--2390. (2018).","journal-title":"Proceedings of the 27th International Conference on Com-putational Linguistics"},{"key":"e_1_3_2_1_32_1","doi-asserted-by":"crossref","unstructured":"Paul Pu Liang Ziyin Liu Amir Zadeh and Louis-Philippe Morency. 2018. \"Multimodal language analysis with recurrent multistage fusion.\" arXiv pre-print arXiv:1808.03920 (2018).  Paul Pu Liang Ziyin Liu Amir Zadeh and Louis-Philippe Morency. 2018. \"Multimodal language analysis with recurrent multistage fusion.\" arXiv pre-print arXiv:1808.03920 (2018).","DOI":"10.18653\/v1\/D18-1014"}],"event":{"name":"MM '19: The 27th ACM International Conference on Multimedia","location":"Nice France","acronym":"MM '19","sponsor":["SIGMM ACM Special Interest Group on Multimedia"]},"container-title":["Proceedings of the 27th ACM International Conference on Multimedia"],"original-title":[],"link":[{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/3343031.3351039","content-type":"unspecified","content-version":"vor","intended-application":"text-mining"},{"URL":"https:\/\/dl.acm.org\/doi\/pdf\/10.1145\/3343031.3351039","content-type":"unspecified","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2025,6,17]],"date-time":"2025-06-17T23:13:11Z","timestamp":1750201991000},"score":1,"resource":{"primary":{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/3343031.3351039"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2019,10,15]]},"references-count":32,"alternative-id":["10.1145\/3343031.3351039","10.1145\/3343031"],"URL":"https:\/\/doi.org\/10.1145\/3343031.3351039","relation":{},"subject":[],"published":{"date-parts":[[2019,10,15]]},"assertion":[{"value":"2019-10-15","order":2,"name":"published","label":"Published","group":{"name":"publication_history","label":"Publication History"}}]}}