{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2025,6,18]],"date-time":"2025-06-18T04:08:52Z","timestamp":1750219732579,"version":"3.41.0"},"publisher-location":"New York, NY, USA","reference-count":30,"publisher":"ACM","license":[{"start":{"date-parts":[[2023,10,29]],"date-time":"2023-10-29T00:00:00Z","timestamp":1698537600000},"content-version":"vor","delay-in-days":0,"URL":"https:\/\/www.acm.org\/publications\/policies\/copyright_policy#Background"}],"funder":[{"name":"Sinergia"}],"content-domain":{"domain":["dl.acm.org"],"crossmark-restriction":true},"short-container-title":[],"published-print":{"date-parts":[[2023,11,2]]},"DOI":"10.1145\/3607542.3617355","type":"proceedings-article","created":{"date-parts":[[2023,10,25]],"date-time":"2023-10-25T00:07:32Z","timestamp":1698192452000},"page":"5-12","update-policy":"https:\/\/doi.org\/10.1145\/crossmark-policy","source":"Crossref","is-referenced-by-count":1,"title":["Latent Wander: an Alternative Interface for Interactive and Serendipitous Discovery of Large AV Archives"],"prefix":"10.1145","author":[{"ORCID":"https:\/\/orcid.org\/0000-0002-6866-1409","authenticated-orcid":false,"given":"Yuchen","family":"Yang","sequence":"first","affiliation":[{"name":"EPFL, Lausanne, Switzerland"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"ORCID":"https:\/\/orcid.org\/0009-0001-5966-5790","authenticated-orcid":false,"given":"Linyida","family":"Zhang","sequence":"additional","affiliation":[{"name":"EPFL, Lausanne, Switzerland"}],"role":[{"role":"author","vocabulary":"crossref"}]}],"member":"320","published-online":{"date-parts":[[2023,10,29]]},"reference":[{"key":"e_1_3_2_1_1_1","first-page":"35","article-title":"A fuzzy inference system for synergy estimation of simultaneous emotion dynamics in agents","volume":"2","author":"Athar Atifa","year":"2011","unstructured":"Atifa Athar , M Saleem Khan , Khalil Ahmed , Aiesha Ahmed , and Nida Anwar . 2011 . A fuzzy inference system for synergy estimation of simultaneous emotion dynamics in agents . Int. J. Sci. Eng. Res 2 , 6 (2011), 35 -- 41 . Atifa Athar, M Saleem Khan, Khalil Ahmed, Aiesha Ahmed, and Nida Anwar. 2011. A fuzzy inference system for synergy estimation of simultaneous emotion dynamics in agents. Int. J. Sci. Eng. Res 2, 6 (2011), 35--41.","journal-title":"Int. J. Sci. Eng. Res"},{"key":"e_1_3_2_1_2_1","doi-asserted-by":"publisher","DOI":"10.1109\/ICCV48922.2021.00175"},{"key":"e_1_3_2_1_3_1","volume-title":"Jeffrey Shaw, and Peter Weibel.","author":"Brown Neil","year":"2000","unstructured":"Neil Brown , Dennis Del Favero , Jeffrey Shaw, and Peter Weibel. 2000 . Interactive narrative as a multi-temporal agency. representations 84 (2000), 88. Neil Brown, Dennis Del Favero, Jeffrey Shaw, and Peter Weibel. 2000. Interactive narrative as a multi-temporal agency. representations 84 (2000), 88."},{"key":"e_1_3_2_1_4_1","doi-asserted-by":"publisher","DOI":"10.1109\/ACIIW52867.2021.9666309"},{"key":"e_1_3_2_1_5_1","volume-title":"The MSR-video to text dataset with clean annotations. arXiv preprint arXiv:2102.06448","author":"Chen Haoran","year":"2021","unstructured":"Haoran Chen , Jianmin Li , Simone Frintrop , and Xiaolin Hu. 2021. The MSR-video to text dataset with clean annotations. arXiv preprint arXiv:2102.06448 ( 2021 ). Haoran Chen, Jianmin Li, Simone Frintrop, and Xiaolin Hu. 2021. The MSR-video to text dataset with clean annotations. arXiv preprint arXiv:2102.06448 (2021)."},{"key":"e_1_3_2_1_6_1","volume-title":"International Conference on Machine Learning. PMLR","author":"Choi Kristy","year":"2020","unstructured":"Kristy Choi , Curtis Hawthorne , Ian Simon , Monica Dinculescu , and Jesse Engel . 2020 . Encoding musical style with transformer autoencoders . In International Conference on Machine Learning. PMLR , 1899--1908. Kristy Choi, Curtis Hawthorne, Ian Simon, Monica Dinculescu, and Jesse Engel. 2020. Encoding musical style with transformer autoencoders. In International Conference on Machine Learning. PMLR, 1899--1908."},{"key":"e_1_3_2_1_7_1","volume-title":"Parrot: Paraphrase generation for NLU. v1. 0","author":"Damodaran Prithiviraj","year":"2021","unstructured":"Prithiviraj Damodaran . 2021 . Parrot: Paraphrase generation for NLU. v1. 0 (2021). Prithiviraj Damodaran. 2021. Parrot: Paraphrase generation for NLU. v1. 0 (2021)."},{"key":"e_1_3_2_1_8_1","volume-title":"An argument for basic emotions. Cognition & emotion 6, 3--4","author":"Ekman Paul","year":"1992","unstructured":"Paul Ekman . 1992. An argument for basic emotions. Cognition & emotion 6, 3--4 ( 1992 ), 169--200. Paul Ekman. 1992. An argument for basic emotions. Cognition & emotion 6, 3--4 (1992), 169--200."},{"key":"e_1_3_2_1_9_1","volume-title":"Proceedings, Part IV 16","author":"Gabeur Valentin","year":"2020","unstructured":"Valentin Gabeur , Chen Sun , Karteek Alahari , and Cordelia Schmid . 2020 . Multimodal transformer for video retrieval. In Computer Vision--ECCV 2020: 16th European Conference, Glasgow, UK, August 23--28, 2020 , Proceedings, Part IV 16 . Springer, 214--229. Valentin Gabeur, Chen Sun, Karteek Alahari, and Cordelia Schmid. 2020. Multimodal transformer for video retrieval. In Computer Vision--ECCV 2020: 16th European Conference, Glasgow, UK, August 23--28, 2020, Proceedings, Part IV 16. Springer, 214--229."},{"key":"e_1_3_2_1_10_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR.2016.90"},{"volume-title":"Museum activism","author":"Janes Robert R","key":"e_1_3_2_1_11_1","unstructured":"Robert R Janes and Richard Sandell . 2019. Museum activism . Taylor & Francis . Robert R Janes and Richard Sandell. 2019. Museum activism. Taylor & Francis."},{"key":"e_1_3_2_1_12_1","doi-asserted-by":"publisher","DOI":"10.1093\/screen\/hjaa057"},{"key":"e_1_3_2_1_13_1","unstructured":"Tamara Klopper. 2022. New installation Film Catcher opened in Eye. https:\/\/www.eyefilm.nl\/nl\/magazine\/nieuwe-installatie-film-catcher-geopendin- eye\/852512  Tamara Klopper. 2022. New installation Film Catcher opened in Eye. https:\/\/www.eyefilm.nl\/nl\/magazine\/nieuwe-installatie-film-catcher-geopendin- eye\/852512"},{"key":"e_1_3_2_1_14_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR46437.2021.00725"},{"key":"e_1_3_2_1_15_1","volume-title":"Taking an Emotional Look at Video Paragraph Captioning. arXiv preprint arXiv:2203.06356","author":"Li Qinyu","year":"2022","unstructured":"Qinyu Li , Tengpeng Li , Hanli Wang , and Chang Wen Chen . 2022. Taking an Emotional Look at Video Paragraph Captioning. arXiv preprint arXiv:2203.06356 ( 2022 ). Qinyu Li, Tengpeng Li, Hanli Wang, and Chang Wen Chen. 2022. Taking an Emotional Look at Video Paragraph Captioning. arXiv preprint arXiv:2203.06356 (2022)."},{"key":"e_1_3_2_1_16_1","volume-title":"Proceedings, Part XIV. Springer, 319--335","author":"Liu Yuqi","year":"2022","unstructured":"Yuqi Liu , Pengfei Xiong , Luhui Xu , Shengming Cao , and Qin Jin . 2022 . Ts2- net: Token shift and selection transformer for text-video retrieval. In Computer Vision--ECCV 2022: 17th European Conference, Tel Aviv, Israel, October 23--27, 2022 , Proceedings, Part XIV. Springer, 319--335 . Yuqi Liu, Pengfei Xiong, Luhui Xu, Shengming Cao, and Qin Jin. 2022. Ts2- net: Token shift and selection transformer for text-video retrieval. In Computer Vision--ECCV 2022: 17th European Conference, Tel Aviv, Israel, October 23--27, 2022, Proceedings, Part XIV. Springer, 319--335."},{"key":"e_1_3_2_1_17_1","doi-asserted-by":"publisher","DOI":"10.1016\/j.neucom.2022.07.028"},{"key":"e_1_3_2_1_18_1","volume-title":"Nanne van Noord, and Giovanna Fossati.","author":"Masson Eef","year":"2020","unstructured":"Eef Masson , Christian Gosvig Olesen , Nanne van Noord, and Giovanna Fossati. 2020 . Exploring Digitised Moving Image Collections: The SEMIA Project, Visual Analysis and the Turn to Abstraction . DHQ : Digital Humanities Quarterly 4 (2020). Eef Masson, Christian Gosvig Olesen, Nanne van Noord, and Giovanna Fossati. 2020. Exploring Digitised Moving Image Collections: The SEMIA Project, Visual Analysis and the Turn to Abstraction. DHQ: Digital Humanities Quarterly 4 (2020)."},{"key":"e_1_3_2_1_19_1","volume-title":"Umap: Uniform manifold approximation and projection for dimension reduction. arXiv preprint arXiv:1802.03426","author":"McInnes Leland","year":"2018","unstructured":"Leland McInnes , John Healy , and James Melville . 2018 . Umap: Uniform manifold approximation and projection for dimension reduction. arXiv preprint arXiv:1802.03426 (2018). Leland McInnes, John Healy, and James Melville. 2018. Umap: Uniform manifold approximation and projection for dimension reduction. arXiv preprint arXiv:1802.03426 (2018)."},{"key":"e_1_3_2_1_20_1","doi-asserted-by":"publisher","DOI":"10.1511\/2001.4.344"},{"key":"e_1_3_2_1_21_1","doi-asserted-by":"publisher","DOI":"10.1075\/idj.22013.rod"},{"key":"e_1_3_2_1_22_1","volume-title":"Transnet V2: an effective deep network architecture for fast shot transition detection. arXiv preprint arXiv:2008.04838","author":"Jakub","year":"2020","unstructured":"Tom\u00e1? Sou?ek and Jakub Loko?. 2020. Transnet V2: an effective deep network architecture for fast shot transition detection. arXiv preprint arXiv:2008.04838 ( 2020 ). Tom\u00e1? Sou?ek and Jakub Loko?. 2020. Transnet V2: an effective deep network architecture for fast shot transition detection. arXiv preprint arXiv:2008.04838 (2020)."},{"key":"e_1_3_2_1_23_1","doi-asserted-by":"publisher","DOI":"10.1109\/ICCV.2019.00756"},{"key":"e_1_3_2_1_24_1","volume-title":"TVLT: Textless Vision-Language Transformer. arXiv preprint arXiv:2209.14156","author":"Tang Zineng","year":"2022","unstructured":"Zineng Tang , Jaemin Cho , Yixin Nie , and Mohit Bansal . 2022 . TVLT: Textless Vision-Language Transformer. arXiv preprint arXiv:2209.14156 (2022). Zineng Tang, Jaemin Cho, Yixin Nie, and Mohit Bansal. 2022. TVLT: Textless Vision-Language Transformer. arXiv preprint arXiv:2209.14156 (2022)."},{"key":"e_1_3_2_1_25_1","volume-title":"Zero-shot video captioning with evolving pseudo-tokens. arXiv preprint arXiv:2207.11100","author":"Tewel Yoad","year":"2022","unstructured":"Yoad Tewel , Yoav Shalev , Roy Nadler , Idan Schwartz , and Lior Wolf . 2022. Zero-shot video captioning with evolving pseudo-tokens. arXiv preprint arXiv:2207.11100 ( 2022 ). Yoad Tewel, Yoav Shalev, Roy Nadler, Idan Schwartz, and Lior Wolf. 2022. Zero-shot video captioning with evolving pseudo-tokens. arXiv preprint arXiv:2207.11100 (2022)."},{"key":"e_1_3_2_1_26_1","first-page":"715","article-title":"Emotion expression with fact transfer for video description","volume":"24","author":"Tang Pengjie","year":"2021","unstructured":"HanliWang, Pengjie Tang , Qinyu Li , and Meng Cheng . 2021 . Emotion expression with fact transfer for video description . IEEE Transactions on Multimedia 24 (2021), 715 -- 727 . HanliWang, Pengjie Tang, Qinyu Li, and Meng Cheng. 2021. Emotion expression with fact transfer for video description. IEEE Transactions on Multimedia 24 (2021), 715--727.","journal-title":"IEEE Transactions on Multimedia"},{"key":"e_1_3_2_1_27_1","doi-asserted-by":"publisher","DOI":"10.1109\/ICCV48922.2021.00677"},{"key":"e_1_3_2_1_28_1","volume-title":"Tableformer: Robust transformer modeling for table-text encoding. arXiv preprint arXiv:2203.00274","author":"Yang Jingfeng","year":"2022","unstructured":"Jingfeng Yang , Aditya Gupta , Shyam Upadhyay , Luheng He , Rahul Goel , and Shachi Paul . 2022 . Tableformer: Robust transformer modeling for table-text encoding. arXiv preprint arXiv:2203.00274 (2022). Jingfeng Yang, Aditya Gupta, Shyam Upadhyay, Luheng He, Rahul Goel, and Shachi Paul. 2022. Tableformer: Robust transformer modeling for table-text encoding. arXiv preprint arXiv:2203.00274 (2022)."},{"key":"e_1_3_2_1_29_1","doi-asserted-by":"publisher","DOI":"10.18653\/v1\/P18-1208"},{"key":"e_1_3_2_1_30_1","doi-asserted-by":"publisher","DOI":"10.1109\/ICCV48922.2021.00299"}],"event":{"name":"MM '23: The 31st ACM International Conference on Multimedia","sponsor":["SIGMM ACM Special Interest Group on Multimedia"],"location":"Ottawa ON Canada","acronym":"MM '23"},"container-title":["Proceedings of the 5th Workshop on analySis, Understanding and proMotion of heritAge Contents"],"original-title":[],"link":[{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/3607542.3617355","content-type":"unspecified","content-version":"vor","intended-application":"text-mining"},{"URL":"https:\/\/dl.acm.org\/doi\/pdf\/10.1145\/3607542.3617355","content-type":"unspecified","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2025,6,17]],"date-time":"2025-06-17T16:36:28Z","timestamp":1750178188000},"score":1,"resource":{"primary":{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/3607542.3617355"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2023,10,29]]},"references-count":30,"alternative-id":["10.1145\/3607542.3617355","10.1145\/3607542"],"URL":"https:\/\/doi.org\/10.1145\/3607542.3617355","relation":{},"subject":[],"published":{"date-parts":[[2023,10,29]]},"assertion":[{"value":"2023-10-29","order":2,"name":"published","label":"Published","group":{"name":"publication_history","label":"Publication History"}}]}}