{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2026,8,2]],"date-time":"2026-08-02T18:06:56Z","timestamp":1785694016175,"version":"3.56.0"},"publisher-location":"New York, NY, USA","reference-count":45,"publisher":"ACM","license":[{"start":{"date-parts":[[2020,8,20]],"date-time":"2020-08-20T00:00:00Z","timestamp":1597881600000},"content-version":"vor","delay-in-days":0,"URL":"https:\/\/www.acm.org\/publications\/policies\/copyright_policy#Background"}],"funder":[{"name":"NSFC","award":["61625107"],"award-info":[{"award-number":["61625107"]}]},{"name":"Zhejiang Natural Science Foundation","award":["LR19F020006"],"award-info":[{"award-number":["LR19F020006"]}]},{"DOI":"10.13039\/501100012226","name":"Fundamental Research Funds for the Central Universities","doi-asserted-by":"publisher","award":["2020QNA5024,2018AAA0101900"],"award-info":[{"award-number":["2020QNA5024,2018AAA0101900"]}],"id":[{"id":"10.13039\/501100012226","id-type":"DOI","asserted-by":"publisher"}]}],"content-domain":{"domain":["dl.acm.org"],"crossmark-restriction":true},"short-container-title":[],"published-print":{"date-parts":[[2020,8,23]]},"DOI":"10.1145\/3394486.3403325","type":"proceedings-article","created":{"date-parts":[[2020,8,20]],"date-time":"2020-08-20T23:03:55Z","timestamp":1597964635000},"page":"2744-2754","update-policy":"https:\/\/doi.org\/10.1145\/crossmark-policy","source":"Crossref","is-referenced-by-count":24,"title":["Comprehensive Information Integration Modeling Framework for Video Titling"],"prefix":"10.1145","author":[{"given":"Shengyu","family":"Zhang","sequence":"first","affiliation":[{"name":"Zhejiang University, Hang Zhou, China"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Ziqi","family":"Tan","sequence":"additional","affiliation":[{"name":"Zhejiang University, Hang Zhou, China"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Zhou","family":"Zhao","sequence":"additional","affiliation":[{"name":"Zhejiang University, Hang Zhou, China"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Jin","family":"Yu","sequence":"additional","affiliation":[{"name":"Alibaba Group, Bei Jing, China"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Kun","family":"Kuang","sequence":"additional","affiliation":[{"name":"Zhejiang University, Hang Zhou, China"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Tan","family":"Jiang","sequence":"additional","affiliation":[{"name":"Zhejiang University, Hang Zhou, China"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Jingren","family":"Zhou","sequence":"additional","affiliation":[{"name":"Alibaba Group, Hang Zhou, China"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Hongxia","family":"Yang","sequence":"additional","affiliation":[{"name":"Alibaba Group, Hang Zhou, China"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Fei","family":"Wu","sequence":"additional","affiliation":[{"name":"Zhejiang University, Hang Zhou, China"}],"role":[{"vocabulary":"crossref","role":"author"}]}],"member":"320","published-online":{"date-parts":[[2020,8,20]]},"reference":[{"key":"e_1_3_2_2_1_1","unstructured":"Dzmitry Bahdanau Kyunghyun Cho and Yoshua Bengio. 2015. Neural Machine Translation by Jointly Learning to Align and Translate.. In ICLR.  Dzmitry Bahdanau Kyunghyun Cho and Yoshua Bengio. 2015. Neural Machine Translation by Jointly Learning to Align and Translate.. In ICLR."},{"key":"e_1_3_2_2_2_1","volume-title":"METEOR: An automatic metric for MT evaluation with improved correlation with human judgments. In ACL. 65--72.","author":"Banerjee Satanjeev","year":"2005","unstructured":"Satanjeev Banerjee and Alon Lavie . 2005 . METEOR: An automatic metric for MT evaluation with improved correlation with human judgments. In ACL. 65--72. Satanjeev Banerjee and Alon Lavie. 2005. METEOR: An automatic metric for MT evaluation with improved correlation with human judgments. In ACL. 65--72."},{"key":"e_1_3_2_2_3_1","volume-title":"Dolan","author":"Chen David L.","year":"2011","unstructured":"David L. Chen and William B . Dolan . 2011 . Collecting Highly Parallel Data for Paraphrase Evaluation.. In ACL. David L. Chen and William B. Dolan. 2011. Collecting Highly Parallel Data for Paraphrase Evaluation.. In ACL."},{"key":"e_1_3_2_2_4_1","doi-asserted-by":"crossref","unstructured":"Mia Xu Chen Orhan Firat Ankur Bapna Melvin Johnson Wolfgang Macherey George Foster Llion Jones Mike Schuster Noam Shazeer Niki Parmar and etal 2018. The Best of Both Worlds: Combining Recent Advances in Neural Machine Translation.. In ACL.  Mia Xu Chen Orhan Firat Ankur Bapna Melvin Johnson Wolfgang Macherey George Foster Llion Jones Mike Schuster Noam Shazeer Niki Parmar and et al. 2018. The Best of Both Worlds: Combining Recent Advances in Neural Machine Translation.. In ACL.","DOI":"10.18653\/v1\/P18-1008"},{"key":"e_1_3_2_2_5_1","doi-asserted-by":"crossref","unstructured":"Qibin Chen Junyang Lin Yichang Zhang Hongxia Yang Jingren Zhou and Jie Tang. 2019. Towards Knowledge-Based Personalized Product Description Generation in E-commerce.. In KDD.  Qibin Chen Junyang Lin Yichang Zhang Hongxia Yang Jingren Zhou and Jie Tang. 2019. Towards Knowledge-Based Personalized Product Description Generation in E-commerce.. In KDD.","DOI":"10.1145\/3292500.3330725"},{"key":"e_1_3_2_2_6_1","doi-asserted-by":"publisher","DOI":"10.1109\/ICMEW.2019.00118"},{"key":"e_1_3_2_2_7_1","volume-title":"Dzmitry Bahdanau, and Yoshua Bengio.","author":"Cho Kyunghyun","year":"2014","unstructured":"Kyunghyun Cho , Bart Van Merri\u00ebnboer , Dzmitry Bahdanau, and Yoshua Bengio. 2014 . On the properties of neural machine translation: Encoder-decoder approaches. arXiv preprint arXiv:1409.1259 (2014). Kyunghyun Cho, Bart Van Merri\u00ebnboer, Dzmitry Bahdanau, and Yoshua Bengio. 2014. On the properties of neural machine translation: Encoder-decoder approaches. arXiv preprint arXiv:1409.1259 (2014)."},{"key":"e_1_3_2_2_8_1","unstructured":"Abhishek Das Satwik Kottur Khushi Gupta Avi Singh Deshraj Yadav Jos\u00e9 M. F. Moura Devi Parikh and Dhruv Batra. 2017. Visual Dialog.. In CVPR.  Abhishek Das Satwik Kottur Khushi Gupta Avi Singh Deshraj Yadav Jos\u00e9 M. F. Moura Devi Parikh and Dhruv Batra. 2017. Visual Dialog.. In CVPR."},{"key":"e_1_3_2_2_9_1","volume":"201","author":"Das Pradipto","unstructured":"Pradipto Das , Chenliang Xu , Richard F. Doell , and Jason J. Corso. 201 3. A Thousand Frames in Just a Few Words: Lingual Description of Videos through Latent Topics and Sparse Object Stitching.. In CVPR. Pradipto Das, Chenliang Xu, Richard F. Doell, and Jason J. Corso. 2013. A Thousand Frames in Just a Few Words: Lingual Description of Videos through Latent Topics and Sparse Object Stitching.. In CVPR.","journal-title":"Jason J. Corso."},{"key":"e_1_3_2_2_10_1","volume-title":"Wildes","author":"Feichtenhofer Christoph","year":"2016","unstructured":"Christoph Feichtenhofer , Axel Pinz , and Richard P . Wildes . 2016 . Spatiotemporal Residual Networks for Video Action Recognition.. In NIPS. Christoph Feichtenhofer, Axel Pinz, and Richard P. Wildes. 2016. Spatiotemporal Residual Networks for Video Action Recognition.. In NIPS."},{"key":"e_1_3_2_2_11_1","volume-title":"ICLR Workshop.","author":"Fey Matthias","year":"2019","unstructured":"Matthias Fey and Jan Eric Lenssen . 2019 . Fast Graph Representation Learning with PyTorch Geometric .. In ICLR Workshop. Matthias Fey and Jan Eric Lenssen. 2019. Fast Graph Representation Learning with PyTorch Geometric.. In ICLR Workshop."},{"key":"e_1_3_2_2_12_1","unstructured":"Jonas Gehring Michael Auli David Grangier Denis Yarats and Yann N Dauphin. 2017. Convolutional sequence to sequence learning. In ICML.  Jonas Gehring Michael Auli David Grangier Denis Yarats and Yann N Dauphin. 2017. Convolutional sequence to sequence learning. In ICML."},{"key":"e_1_3_2_2_13_1","doi-asserted-by":"crossref","unstructured":"Sergio Guadarrama Niveda Krishnamoorthy Girish Malkarnenkar Subhashini Venugopalan Raymond J. Mooney Trevor Darrell and Kate Saenko. 2013. YouTube2Text: Recognizing and Describing Arbitrary Activities Using Semantic Hierarchies and Zero-Shot Recognition.. In ICCV.  Sergio Guadarrama Niveda Krishnamoorthy Girish Malkarnenkar Subhashini Venugopalan Raymond J. Mooney Trevor Darrell and Kate Saenko. 2013. YouTube2Text: Recognizing and Describing Arbitrary Activities Using Semantic Hierarchies and Zero-Shot Recognition.. In ICCV.","DOI":"10.1109\/ICCV.2013.337"},{"key":"e_1_3_2_2_14_1","doi-asserted-by":"crossref","unstructured":"Kensho Hara Hirokatsu Kataoka and Yutaka Satoh. 2018. Can Spatiotemporal 3D CNNs Retrace the History of 2D CNNs and ImageNet?. In CVPR.  Kensho Hara Hirokatsu Kataoka and Yutaka Satoh. 2018. Can Spatiotemporal 3D CNNs Retrace the History of 2D CNNs and ImageNet?. In CVPR.","DOI":"10.1109\/CVPR.2018.00685"},{"key":"e_1_3_2_2_15_1","volume-title":"Kingma and Jimmy Ba","author":"Diederik","year":"2015","unstructured":"Diederik P. Kingma and Jimmy Ba . 2015 . Adam : A Method for Stochastic Optimization.. In ICML. Diederik P. Kingma and Jimmy Ba. 2015. Adam: A Method for Stochastic Optimization.. In ICML."},{"key":"e_1_3_2_2_16_1","unstructured":"Thomas N Kipf and Max Welling. 2017. Semi-supervised classification with graph convolutional networks. In ICLR.  Thomas N Kipf and Max Welling. 2017. Semi-supervised classification with graph convolutional networks. In ICLR."},{"key":"e_1_3_2_2_17_1","volume-title":"Natural Language Description of Human Activities from Video Images Based on Concept Hierarchy of Actions. IJCV","author":"Kojima Atsuhiro","year":"2002","unstructured":"Atsuhiro Kojima , Takeshi Tamura , and Kunio Fukunaga . 2002. Natural Language Description of Human Activities from Video Images Based on Concept Hierarchy of Actions. IJCV ( 2002 ). Atsuhiro Kojima, Takeshi Tamura, and Kunio Fukunaga. 2002. Natural Language Description of Human Activities from Video Images Based on Concept Hierarchy of Actions. IJCV (2002)."},{"key":"e_1_3_2_2_18_1","unstructured":"Jiwei Li Will Monroe Tianlin Shi S\u00e9bastien Jean Alan Ritter and Dan Jurafsky. 2017. Adversarial Learning for Neural Dialogue Generation.. In EMNLP.  Jiwei Li Will Monroe Tianlin Shi S\u00e9bastien Jean Alan Ritter and Dan Jurafsky. 2017. Adversarial Learning for Neural Dialogue Generation.. In EMNLP."},{"key":"e_1_3_2_2_19_1","volume-title":"Rouge: A package for automatic evaluation of summaries. In ACL. 74--81.","author":"Lin Chin-Yew","year":"2004","unstructured":"Chin-Yew Lin . 2004 . Rouge: A package for automatic evaluation of summaries. In ACL. 74--81. Chin-Yew Lin. 2004. Rouge: A package for automatic evaluation of summaries. In ACL. 74--81."},{"key":"e_1_3_2_2_20_1","volume-title":"Deep Fashion Analysis with Feature Map Upsampling and Landmark-Driven Attention","author":"Liu Jingyuan","unstructured":"Jingyuan Liu and Hong Lu. 2018. Deep Fashion Analysis with Feature Map Upsampling and Landmark-Driven Attention . In ECCV. Springer . Jingyuan Liu and Hong Lu. 2018. Deep Fashion Analysis with Feature Map Upsampling and Landmark-Driven Attention. In ECCV. Springer."},{"key":"e_1_3_2_2_21_1","doi-asserted-by":"crossref","unstructured":"Ziwei Liu Sijie Yan Ping Luo Xiaogang Wang and Xiaoou Tang. 2016. Fashion Landmark Detection in the Wild.. In ECCV.  Ziwei Liu Sijie Yan Ping Luo Xiaogang Wang and Xiaoou Tang. 2016. Fashion Landmark Detection in the Wild.. In ECCV.","DOI":"10.1007\/978-3-319-46475-6_15"},{"key":"e_1_3_2_2_22_1","volume-title":"Vilbert: Pretraining task-agnostic visiolinguistic representations for vision-and-language tasks. In NIPS.","author":"Lu Jiasen","year":"2019","unstructured":"Jiasen Lu , Dhruv Batra , Devi Parikh , and Stefan Lee . 2019 . Vilbert: Pretraining task-agnostic visiolinguistic representations for vision-and-language tasks. In NIPS. Jiasen Lu, Dhruv Batra, Devi Parikh, and Stefan Lee. 2019. Vilbert: Pretraining task-agnostic visiolinguistic representations for vision-and-language tasks. In NIPS."},{"key":"e_1_3_2_2_23_1","unstructured":"Shuming Ma Lei Cui Damai Dai Furu Wei and Xu Sun. 2019. LiveBot: Generating Live Video Comments Based on Visual and Textual Contexts.. In AAAI.  Shuming Ma Lei Cui Damai Dai Furu Wei and Xu Sun. 2019. LiveBot: Generating Live Video Comments Based on Visual and Textual Contexts.. In AAAI."},{"key":"e_1_3_2_2_24_1","unstructured":"Pingbo Pan Zhongwen Xu Yi Yang Fei Wu and Yueting Zhuang. 2016. Hierarchical Recurrent Neural Encoder for Video Representation with Application to Captioning.. In CVPR.  Pingbo Pan Zhongwen Xu Yi Yang Fei Wu and Yueting Zhuang. 2016. Hierarchical Recurrent Neural Encoder for Video Representation with Application to Captioning.. In CVPR."},{"key":"e_1_3_2_2_25_1","volume-title":"BLEU: a method for automatic evaluation of machine translation","author":"Papineni Kishore","unstructured":"Kishore Papineni , Salim Roukos , Todd Ward , and Wei-Jing Zhu . 2002. BLEU: a method for automatic evaluation of machine translation . In ACL. Association for Computational Linguistics , 311--318. Kishore Papineni, Salim Roukos, Todd Ward, and Wei-Jing Zhu. 2002. BLEU: a method for automatic evaluation of machine translation. In ACL. Association for Computational Linguistics, 311--318."},{"key":"e_1_3_2_2_26_1","unstructured":"Adam Paszke Sam Gross Soumith Chintala Gregory Chanan Edward Yang Zachary DeVito Zeming Lin Alban Desmaison Luca Antiga and Adam Lerer. 2017. Automatic differentiation in PyTorch. In NIPS-W.  Adam Paszke Sam Gross Soumith Chintala Gregory Chanan Edward Yang Zachary DeVito Zeming Lin Alban Desmaison Luca Antiga and Adam Lerer. 2017. Automatic differentiation in PyTorch. In NIPS-W."},{"key":"e_1_3_2_2_27_1","doi-asserted-by":"crossref","unstructured":"Xufeng Qian Yueting Zhuang Yimeng Li Shaoning Xiao Shiliang Pu and Jun Xiao. 2019. Video Relation Detection with Spatio-Temporal Graph.. In MM.  Xufeng Qian Yueting Zhuang Yimeng Li Shaoning Xiao Shiliang Pu and Jun Xiao. 2019. Video Relation Detection with Spatio-Temporal Graph.. In MM.","DOI":"10.1145\/3343031.3351058"},{"key":"e_1_3_2_2_28_1","volume-title":"Grounding Action Descriptions in Videos. TACL","author":"Regneri Michaela","year":"2013","unstructured":"Michaela Regneri , Marcus Rohrbach , Dominikus Wetzel , Stefan Thater , Bernt Schiele , and Manfred Pinkal . 2013. Grounding Action Descriptions in Videos. TACL ( 2013 ). Michaela Regneri, Marcus Rohrbach, Dominikus Wetzel, Stefan Thater, Bernt Schiele, and Manfred Pinkal. 2013. Grounding Action Descriptions in Videos. TACL (2013)."},{"key":"e_1_3_2_2_29_1","doi-asserted-by":"crossref","unstructured":"Anna Rohrbach Marcus Rohrbach Wei Qiu Annemarie Friedrich Manfred Pinkal and Bernt Schiele. 2014. Coherent Multi-sentence Video Description with Variable Level of Detail.. In GCPR.  Anna Rohrbach Marcus Rohrbach Wei Qiu Annemarie Friedrich Manfred Pinkal and Bernt Schiele. 2014. Coherent Multi-sentence Video Description with Variable Level of Detail.. In GCPR.","DOI":"10.1007\/978-3-319-11752-2_15"},{"key":"e_1_3_2_2_30_1","doi-asserted-by":"crossref","unstructured":"Anna Rohrbach Marcus Rohrbach Niket Tandon and Bernt Schiele. 2015. A dataset for Movie Description.. In CVPR.  Anna Rohrbach Marcus Rohrbach Niket Tandon and Bernt Schiele. 2015. A dataset for Movie Description.. In CVPR.","DOI":"10.1109\/CVPR.2015.7298940"},{"key":"e_1_3_2_2_31_1","doi-asserted-by":"crossref","unstructured":"Marcus Rohrbach Wei Qiu Ivan Titov Stefan Thater Manfred Pinkal and Bernt Schiele. 2013. Translating Video Content to Natural Language Descriptions.. In ICCV.  Marcus Rohrbach Wei Qiu Ivan Titov Stefan Thater Manfred Pinkal and Bernt Schiele. 2013. Translating Video Content to Natural Language Descriptions.. In ICCV.","DOI":"10.1109\/ICCV.2013.61"},{"key":"e_1_3_2_2_32_1","volume-title":"Manning","author":"Liu Peter J.","year":"2017","unstructured":"Abigail See, Peter J. Liu , and Christopher D . Manning . 2017 . Get To The Point: Summarization with Pointer-Generator Networks.. In ACL. Abigail See, Peter J. Liu, and Christopher D. Manning. 2017. Get To The Point: Summarization with Pointer-Generator Networks.. In ACL."},{"key":"e_1_3_2_2_33_1","doi-asserted-by":"crossref","unstructured":"Gunnar A. Sigurdsson G\u00fcl Varol Xiaolong Wang Ali Farhadi Ivan Laptev and Abhinav Gupta. 2016. Hollywood in Homes: Crowdsourcing Data Collection for Activity Understanding.. In ECCV.  Gunnar A. Sigurdsson G\u00fcl Varol Xiaolong Wang Ali Farhadi Ivan Laptev and Abhinav Gupta. 2016. Hollywood in Homes: Crowdsourcing Data Collection for Activity Understanding.. In ECCV.","DOI":"10.1007\/978-3-319-46448-0_31"},{"key":"e_1_3_2_2_34_1","volume-title":"Courville","author":"Torabi Atousa","year":"2015","unstructured":"Atousa Torabi , Christopher J. Pal , Hugo Larochelle , and Aaron C . Courville . 2015 . Using Descriptive Video Services to Create a Large Data Source for Video Annotation Research. Arxiv ( 2015). Atousa Torabi, Christopher J. Pal, Hugo Larochelle, and Aaron C. Courville. 2015. Using Descriptive Video Services to Create a Large Data Source for Video Annotation Research. Arxiv (2015)."},{"key":"e_1_3_2_2_35_1","volume-title":"Cider: Consensus-based image description evaluation. In CVPR. 4566--4575.","author":"Vedantam Ramakrishna","year":"2015","unstructured":"Ramakrishna Vedantam , C Lawrence Zitnick , and Devi Parikh . 2015 . Cider: Consensus-based image description evaluation. In CVPR. 4566--4575. Ramakrishna Vedantam, C Lawrence Zitnick, and Devi Parikh. 2015. Cider: Consensus-based image description evaluation. In CVPR. 4566--4575."},{"key":"e_1_3_2_2_36_1","doi-asserted-by":"crossref","unstructured":"Subhashini Venugopalan Marcus Rohrbach Jeffrey Donahue Raymond J. Mooney Trevor Darrell and Kate Saenko. 2015a. Sequence to Sequence - Video to Text.. In ICCV.  Subhashini Venugopalan Marcus Rohrbach Jeffrey Donahue Raymond J. Mooney Trevor Darrell and Kate Saenko. 2015a. Sequence to Sequence - Video to Text.. In ICCV.","DOI":"10.1109\/ICCV.2015.515"},{"key":"e_1_3_2_2_37_1","doi-asserted-by":"crossref","unstructured":"Subhashini Venugopalan Huijuan Xu Jeff Donahue Marcus Rohrbach Raymond J. Mooney and Kate Saenko. 2015b. Translating Videos to Natural Language Using Deep Recurrent Neural Networks.. In NAACL.  Subhashini Venugopalan Huijuan Xu Jeff Donahue Marcus Rohrbach Raymond J. Mooney and Kate Saenko. 2015b. Translating Videos to Natural Language Using Deep Recurrent Neural Networks.. In NAACL.","DOI":"10.3115\/v1\/N15-1173"},{"key":"e_1_3_2_2_38_1","doi-asserted-by":"crossref","unstructured":"Bairui Wang Lin Ma Wei Zhang and Wei Liu. 2018. Reconstruction Network for Video Captioning.. In CVPR.  Bairui Wang Lin Ma Wei Zhang and Wei Liu. 2018. Reconstruction Network for Video Captioning.. In CVPR.","DOI":"10.1109\/CVPR.2018.00795"},{"key":"e_1_3_2_2_39_1","doi-asserted-by":"crossref","unstructured":"Wenguan Wang Xiankai Lu Jianbing Shen David J Crandall and Ling Shao. 2019. Zero-shot video object segmentation via attentive graph neural networks. In ICCV.  Wenguan Wang Xiankai Lu Jianbing Shen David J Crandall and Ling Shao. 2019. Zero-shot video object segmentation via attentive graph neural networks. In ICCV.","DOI":"10.1109\/ICCV.2019.00933"},{"key":"e_1_3_2_2_40_1","doi-asserted-by":"crossref","unstructured":"Xiaolong Wang and Abhinav Gupta. 2018. Videos as Space-Time Region Graphs.. In ECCV.  Xiaolong Wang and Abhinav Gupta. 2018. Videos as Space-Time Region Graphs.. In ECCV.","DOI":"10.1007\/978-3-030-01228-1_25"},{"key":"e_1_3_2_2_41_1","volume-title":"Williams and David Zipser","author":"Ronald","year":"1989","unstructured":"Ronald J. Williams and David Zipser . 1989 . A Learning Algorithm for Continually Running Fully Recurrent Neural Networks. Neural Computation ( 1989). Ronald J. Williams and David Zipser. 1989. A Learning Algorithm for Continually Running Fully Recurrent Neural Networks. Neural Computation (1989)."},{"key":"e_1_3_2_2_42_1","volume-title":"Courville","author":"Yao Li","year":"2015","unstructured":"Li Yao , Atousa Torabi , Kyunghyun Cho , Nicolas Ballas , Christopher J. Pal , Hugo Larochelle , and Aaron C . Courville . 2015 . Describing Videos by Exploiting Temporal Structure.. In ICCV. Li Yao, Atousa Torabi, Kyunghyun Cho, Nicolas Ballas, Christopher J. Pal, Hugo Larochelle, and Aaron C. Courville. 2015. Describing Videos by Exploiting Temporal Structure.. In ICCV."},{"key":"e_1_3_2_2_43_1","doi-asserted-by":"crossref","unstructured":"Daniel Zeman Jan Hajic Martin Popel Martin Potthast Milan Straka Filip Ginter Joakim Nivre and Slav Petrov. 2018. CoNLL 2018 Shared Task: Multilingual Parsing from Raw Text to Universal Dependencies.. In ACL.  Daniel Zeman Jan Hajic Martin Popel Martin Potthast Milan Straka Filip Ginter Joakim Nivre and Slav Petrov. 2018. CoNLL 2018 Shared Task: Multilingual Parsing from Raw Text to Universal Dependencies.. In ACL.","DOI":"10.18653\/v1\/K17-3001"},{"key":"e_1_3_2_2_44_1","volume-title":"Juan Carlos Niebles, and Min Sun","author":"Zeng Kuo-Hao","year":"2016","unstructured":"Kuo-Hao Zeng , Tseng-Hung Chen , Juan Carlos Niebles, and Min Sun . 2016 . Title Generation for User Generated Videos.. In ECCV. Kuo-Hao Zeng, Tseng-Hung Chen, Juan Carlos Niebles, and Min Sun. 2016. Title Generation for User Generated Videos.. In ECCV."},{"key":"e_1_3_2_2_45_1","doi-asserted-by":"crossref","unstructured":"Runhao Zeng Wenbing Huang Mingkui Tan Yu Rong Peilin Zhao Junzhou Huang and Chuang Gan. 2019. Graph convolutional networks for temporal action localization. In ICCV.  Runhao Zeng Wenbing Huang Mingkui Tan Yu Rong Peilin Zhao Junzhou Huang and Chuang Gan. 2019. Graph convolutional networks for temporal action localization. In ICCV.","DOI":"10.1109\/ICCV.2019.00719"}],"event":{"name":"KDD '20: The 26th ACM SIGKDD Conference on Knowledge Discovery and Data Mining","location":"Virtual Event CA USA","acronym":"KDD '20","sponsor":["SIGMOD ACM Special Interest Group on Management of Data","SIGKDD ACM Special Interest Group on Knowledge Discovery in Data"]},"container-title":["Proceedings of the 26th ACM SIGKDD International Conference on Knowledge Discovery &amp; Data Mining"],"original-title":[],"link":[{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/3394486.3403325","content-type":"unspecified","content-version":"vor","intended-application":"text-mining"},{"URL":"https:\/\/dl.acm.org\/doi\/pdf\/10.1145\/3394486.3403325","content-type":"unspecified","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2025,6,17]],"date-time":"2025-06-17T22:01:49Z","timestamp":1750197709000},"score":1,"resource":{"primary":{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/3394486.3403325"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2020,8,20]]},"references-count":45,"alternative-id":["10.1145\/3394486.3403325","10.1145\/3394486"],"URL":"https:\/\/doi.org\/10.1145\/3394486.3403325","relation":{},"subject":[],"published":{"date-parts":[[2020,8,20]]},"assertion":[{"value":"2020-08-20","order":2,"name":"published","label":"Published","group":{"name":"publication_history","label":"Publication History"}}]}}