{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2026,4,26]],"date-time":"2026-04-26T03:49:36Z","timestamp":1777175376139,"version":"3.51.4"},"reference-count":30,"publisher":"Association for Computing Machinery (ACM)","issue":"2","license":[{"start":{"date-parts":[[2019,4,29]],"date-time":"2019-04-29T00:00:00Z","timestamp":1556496000000},"content-version":"vor","delay-in-days":0,"URL":"https:\/\/www.acm.org\/publications\/policies\/copyright_policy#Background"}],"content-domain":{"domain":["dl.acm.org"],"crossmark-restriction":true},"short-container-title":["J. Comput. Cult. Herit."],"published-print":{"date-parts":[[2019,6,21]]},"abstract":"<jats:p>We consider the problem of localizing visitors in a cultural site from egocentric (first-person) images. Localization information can be useful both to assist the user during his visit (e.g., by suggesting where to go and what to see next) and to provide behavioral information to the manager of the cultural site (e.g., how much time has been spent by visitors at a given location? What has been liked most?). To tackle the problem, we collected a large dataset of egocentric videos using two cameras: a head-mounted HoloLens device and a chest-mounted GoPro. Each frame has been labeled according to the location of the visitor and to what he was looking at. The dataset is freely available in order to encourage research in this domain. The dataset is complemented with baseline experiments performed considering a state-of-the-art method for location-based temporal segmentation of egocentric videos. Experiments show that compelling results can be achieved to extract useful information for both the visitor and the site-manager.<\/jats:p>","DOI":"10.1145\/3276772","type":"journal-article","created":{"date-parts":[[2019,4,29]],"date-time":"2019-04-29T17:12:14Z","timestamp":1556557934000},"page":"1-19","update-policy":"https:\/\/doi.org\/10.1145\/crossmark-policy","source":"Crossref","is-referenced-by-count":22,"title":["Egocentric Visitors Localization in Cultural Sites"],"prefix":"10.1145","volume":"12","author":[{"given":"Francesco","family":"Ragusa","sequence":"first","affiliation":[{"name":"DMI - IPLab, Universit\u00e0 degli Studi di Catania"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Antonino","family":"Furnari","sequence":"additional","affiliation":[{"name":"DMI - IPLab, Universit\u00e0 degli Studi di Catania"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Sebastiano","family":"Battiato","sequence":"additional","affiliation":[{"name":"DMI - IPLab, Universit\u00e0 degli Studi di Catania"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Giovanni","family":"Signorello","sequence":"additional","affiliation":[{"name":"CUTGANA, Universit\u00e0 degli Studi di Catania"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Giovanni Maria","family":"Farinella","sequence":"additional","affiliation":[{"name":"DMI - IPLab 8 CUTGANA, Universit\u00e0 degli Studi di Catania"}],"role":[{"role":"author","vocabulary":"crossref"}]}],"member":"320","published-online":{"date-parts":[[2019,4,29]]},"reference":[{"key":"e_1_2_2_1_1","doi-asserted-by":"publisher","DOI":"10.1145\/2724727"},{"key":"e_1_2_2_2_1","volume-title":"Proceedings of the Workshop on Perceptual User Interfaces. ACM, 79--82","author":"Aoki Hisashi","year":"1998","unstructured":"Hisashi Aoki , Bernt Schiele , and Alex Pentland . 1998 . Recognizing personal location from video . In Proceedings of the Workshop on Perceptual User Interfaces. ACM, 79--82 . Hisashi Aoki, Bernt Schiele, and Alex Pentland. 1998. Recognizing personal location from video. In Proceedings of the Workshop on Perceptual User Interfaces. ACM, 79--82."},{"key":"e_1_2_2_3_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPRW.2015.7301279"},{"key":"e_1_2_2_4_1","volume-title":"Pattern Recognition and Machine Learning","author":"Bishop C. M.","unstructured":"C. M. Bishop . 2006. Pattern Recognition and Machine Learning . Springer . C. M. Bishop. 2006. Pattern Recognition and Machine Learning. Springer."},{"key":"e_1_2_2_5_1","doi-asserted-by":"publisher","DOI":"10.1109\/SITIS.2014.14"},{"key":"e_1_2_2_6_1","doi-asserted-by":"publisher","DOI":"10.1109\/MMUL.2014.19"},{"key":"e_1_2_2_7_1","doi-asserted-by":"publisher","DOI":"10.1080\/17489725.2011.562927"},{"key":"e_1_2_2_8_1","volume-title":"Personal-location-based temporal segmentation of egocentric video for lifelogging applications. Journal of Visual Communication and Image Representation","author":"Furnari Antonino","year":"2018","unstructured":"Antonino Furnari , Sebastiano Battiato , and Giovanni Maria Farinella . 2018. Personal-location-based temporal segmentation of egocentric video for lifelogging applications. Journal of Visual Communication and Image Representation ( 2018 ). Antonino Furnari, Sebastiano Battiato, and Giovanni Maria Farinella. 2018. Personal-location-based temporal segmentation of egocentric video for lifelogging applications. Journal of Visual Communication and Image Representation (2018)."},{"key":"e_1_2_2_9_1","volume-title":"Proceedings of the International Conference on Image Analysis and Processing","volume":"10485","author":"Gallo G.","unstructured":"G. Gallo , G. Signorello , G. M. Farinella , and A. Torrisi . 2017. Exploiting social images to understand tourist behaviour . In Proceedings of the International Conference on Image Analysis and Processing , Vol. LNCS 10485 . Springer, 707--717. G. Gallo, G. Signorello, G. M. Farinella, and A. Torrisi. 2017. Exploiting social images to understand tourist behaviour. In Proceedings of the International Conference on Image Analysis and Processing, Vol. LNCS 10485. Springer, 707--717."},{"key":"e_1_2_2_10_1","doi-asserted-by":"publisher","DOI":"10.1109\/SURV.2009.090103"},{"key":"e_1_2_2_11_1","doi-asserted-by":"publisher","DOI":"10.1109\/ICASSP.2017.7953298"},{"key":"e_1_2_2_12_1","doi-asserted-by":"publisher","DOI":"10.1109\/WACV.2017.91"},{"key":"e_1_2_2_13_1","doi-asserted-by":"publisher","DOI":"10.1145\/2647868.2654889"},{"key":"e_1_2_2_14_1","doi-asserted-by":"publisher","DOI":"10.1109\/ICCV.2015.336"},{"key":"e_1_2_2_15_1","doi-asserted-by":"publisher","DOI":"10.1145\/3149808.3149812"},{"key":"e_1_2_2_16_1","doi-asserted-by":"publisher","DOI":"10.1023\/A:1011139631724"},{"key":"e_1_2_2_17_1","doi-asserted-by":"publisher","DOI":"10.1109\/ICCVW.2017.281"},{"key":"e_1_2_2_18_1","unstructured":"Maxime Portaz Johann Poignant Mateusz Budnik Philippe Mulhem Jean-Pierre Chevallet and Lorraine Goeuriot. 2017b. Construction et \u00e9valuation d\u2019un corpus pour la recherche d\u2019instances d\u2019images mus\u00e9ales. In CORIA. 17--34.  Maxime Portaz Johann Poignant Mateusz Budnik Philippe Mulhem Jean-Pierre Chevallet and Lorraine Goeuriot. 2017b. Construction et \u00e9valuation d\u2019un corpus pour la recherche d\u2019instances d\u2019images mus\u00e9ales. In CORIA. 17--34."},{"key":"e_1_2_2_19_1","doi-asserted-by":"publisher","DOI":"10.1109\/TPAMI.2016.2611662"},{"key":"e_1_2_2_20_1","doi-asserted-by":"publisher","DOI":"10.1145\/3092832"},{"key":"e_1_2_2_21_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR.2013.377"},{"key":"e_1_2_2_22_1","doi-asserted-by":"publisher","DOI":"10.1007\/978-3-319-25903-1_72"},{"key":"e_1_2_2_23_1","unstructured":"K. Simonyan and A. Zisserman. 2014. Very deep convolutional networks for large-scale image recognition. CoRR abs\/1409.1556 (2014).  K. Simonyan and A. Zisserman. 2014. Very deep convolutional networks for large-scale image recognition. CoRR abs\/1409.1556 (2014)."},{"key":"e_1_2_2_24_1","volume-title":"Proceedings of the International Symposium on Wearable Computing. 50--57","author":"Starner T.","unstructured":"T. Starner , B. Schiele , and A. Pentland . 1998. Visual contextual awareness in wearable computing . In Proceedings of the International Symposium on Wearable Computing. 50--57 . T. Starner, B. Schiele, and A. Pentland. 1998. Visual contextual awareness in wearable computing. In Proceedings of the International Symposium on Wearable Computing. 50--57."},{"key":"e_1_2_2_25_1","doi-asserted-by":"publisher","DOI":"10.1145\/2964284.2973813"},{"key":"e_1_2_2_26_1","volume-title":"Computer Vision, 2003. Proceedings. Ninth IEEE International Conference on. IEEE, 273--280","author":"Torralba Antonio","unstructured":"Antonio Torralba , Kevin P. Murphy , William T. Freeman , and Mark A. Rubin . 2003. Context-based vision system for place and object recognition . In Computer Vision, 2003. Proceedings. Ninth IEEE International Conference on. IEEE, 273--280 . Antonio Torralba, Kevin P. Murphy, William T. Freeman, and Mark A. Rubin. 2003. Context-based vision system for place and object recognition. In Computer Vision, 2003. Proceedings. Ninth IEEE International Conference on. IEEE, 273--280."},{"key":"e_1_2_2_27_1","doi-asserted-by":"publisher","DOI":"10.1109\/30.125076"},{"key":"e_1_2_2_28_1","doi-asserted-by":"publisher","DOI":"10.1016\/j.cviu.2015.02.002"},{"key":"e_1_2_2_29_1","doi-asserted-by":"publisher","DOI":"10.1145\/2628363.2628390"},{"key":"e_1_2_2_30_1","unstructured":"B. Zhou A. Lapedriza J. Xiao A. Torralba and A. Oliva. 2014. Learning deep features for scene recognition using places database. In Advances in Neural Information Processing Systems. 487--495.   B. Zhou A. Lapedriza J. Xiao A. Torralba and A. Oliva. 2014. Learning deep features for scene recognition using places database. In Advances in Neural Information Processing Systems. 487--495."}],"container-title":["Journal on Computing and Cultural Heritage"],"original-title":[],"language":"en","link":[{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/3276772","content-type":"unspecified","content-version":"vor","intended-application":"text-mining"},{"URL":"https:\/\/dl.acm.org\/doi\/pdf\/10.1145\/3276772","content-type":"unspecified","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2025,6,18]],"date-time":"2025-06-18T19:03:40Z","timestamp":1750273420000},"score":1,"resource":{"primary":{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/3276772"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2019,4,29]]},"references-count":30,"journal-issue":{"issue":"2","published-print":{"date-parts":[[2019,6,21]]}},"alternative-id":["10.1145\/3276772"],"URL":"https:\/\/doi.org\/10.1145\/3276772","relation":{},"ISSN":["1556-4673","1556-4711"],"issn-type":[{"value":"1556-4673","type":"print"},{"value":"1556-4711","type":"electronic"}],"subject":[],"published":{"date-parts":[[2019,4,29]]},"assertion":[{"value":"2018-02-01","order":0,"name":"received","label":"Received","group":{"name":"publication_history","label":"Publication History"}},{"value":"2018-09-01","order":1,"name":"accepted","label":"Accepted","group":{"name":"publication_history","label":"Publication History"}},{"value":"2019-04-29","order":2,"name":"published","label":"Published","group":{"name":"publication_history","label":"Publication History"}}]}}