{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2025,10,31]],"date-time":"2025-10-31T22:03:38Z","timestamp":1761948218145,"version":"3.41.0"},"reference-count":48,"publisher":"Association for Computing Machinery (ACM)","issue":"4","license":[{"start":{"date-parts":[[2016,12,19]],"date-time":"2016-12-19T00:00:00Z","timestamp":1482105600000},"content-version":"vor","delay-in-days":0,"URL":"https:\/\/www.acm.org\/publications\/policies\/copyright_policy#Background"}],"funder":[{"DOI":"10.13039\/501100000780","name":"European Commission","doi-asserted-by":"crossref","award":["CIP-ICT-PSP.2012.2.1"],"award-info":[{"award-number":["CIP-ICT-PSP.2012.2.1"]}],"id":[{"id":"10.13039\/501100000780","id-type":"DOI","asserted-by":"crossref"}]},{"name":"EAGLE the Europeana network of Ancient Greek and Latin Epigraphy"},{"name":"Europeana and creativity","award":["325122"],"award-info":[{"award-number":["325122"]}]}],"content-domain":{"domain":["dl.acm.org"],"crossmark-restriction":true},"short-container-title":["J. Comput. Cult. Herit."],"published-print":{"date-parts":[[2016,12,19]]},"abstract":"<jats:p>By bringing together the most prominent European institutions and archives in the field of Classical Latin and Greek epigraphy, the EAGLE project has collected the vast majority of the surviving Greco-Latin inscriptions into a single readily-searchable database. Text-based search engines are typically used to retrieve information about ancient inscriptions (or about other artifacts). These systems require that the users formulate a text query that contains information such as the place where the object was found or where it is currently located. Conversely, visual search systems can be used to provide information to users (like tourists and scholars) in a most intuitive and immediate way, just using an image as query. In this article, we provide a comparison of several approaches for visual recognizing ancient inscriptions. Our experiments, conducted on 17, 155 photos related to 14, 560 inscriptions, show that BoW and VLAD are outperformed by both Fisher Vector (FV) and Convolutional Neural Network (CNN) features. More interestingly, combining FV and CNN features into a single image representation allows achieving very high effectiveness by correctly recognizing the query inscription in more than 90% of the cases. Our results suggest that combinations of FV and CNN can be also exploited to effectively perform visual retrieval of other types of objects related to cultural heritage such as landmarks and monuments.<\/jats:p>","DOI":"10.1145\/2964911","type":"journal-article","created":{"date-parts":[[2016,12,20]],"date-time":"2016-12-20T13:25:27Z","timestamp":1482240327000},"page":"1-24","update-policy":"https:\/\/doi.org\/10.1145\/crossmark-policy","source":"Crossref","is-referenced-by-count":18,"title":["Visual Recognition of Ancient Inscriptions Using Convolutional Neural Network and Fisher Vector"],"prefix":"10.1145","volume":"9","author":[{"given":"Giuseppe","family":"Amato","sequence":"first","affiliation":[{"name":"ISTI-CNR, Pisa, Italy"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Fabrizio","family":"Falchi","sequence":"additional","affiliation":[{"name":"ISTI-CNR, Pisa, Italy"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Lucia","family":"Vadicamo","sequence":"additional","affiliation":[{"name":"ISTI-CNR, Pisa, Italy"}],"role":[{"role":"author","vocabulary":"crossref"}]}],"member":"320","published-online":{"date-parts":[[2016,12,19]]},"reference":[{"key":"e_1_2_1_1_1","unstructured":"Epigraphic Database Roma. 1999. Retrieved from http:\/\/www.edr-edr.it.  Epigraphic Database Roma. 1999. Retrieved from http:\/\/www.edr-edr.it."},{"volume-title":"Proceedings of the International Conference on Computer Vision Theory and Applications (VISIGRAPP\u201913)","author":"Amato G.","key":"e_1_2_1_2_1","unstructured":"G. Amato , F. Falchi , and C. Gennaro . 2013. On reducing the number of visual words in the bag-of-features representation . In Proceedings of the International Conference on Computer Vision Theory and Applications (VISIGRAPP\u201913) . 657--662. G. Amato, F. Falchi, and C. Gennaro. 2013. On reducing the number of visual words in the bag-of-features representation. In Proceedings of the International Conference on Computer Vision Theory and Applications (VISIGRAPP\u201913). 657--662."},{"key":"e_1_2_1_3_1","doi-asserted-by":"publisher","DOI":"10.1145\/2724727"},{"key":"e_1_2_1_4_1","volume-title":"Proceedings of the 1st EAGLE International Conference","volume":"26","author":"Amato G.","year":"2015","unstructured":"G. Amato , F. Falchi , F. Rabitti , and L. Vadicamo . 2014. Inscriptions visual recognition. A comparison of state-of-the-art object recognition approaches . In Proceedings of the 1st EAGLE International Conference , Vol. 26 . Sapienza Universit\u00e1 Editrice, 117--131. http:\/\/archiv.ub.uni-heidelberg.de\/propylaeumdok\/volltexte\/ 2015 \/2337. G. Amato, F. Falchi, F. Rabitti, and L. Vadicamo. 2014. Inscriptions visual recognition. A comparison of state-of-the-art object recognition approaches. In Proceedings of the 1st EAGLE International Conference, Vol. 26. Sapienza Universit\u00e1 Editrice, 117--131. http:\/\/archiv.ub.uni-heidelberg.de\/propylaeumdok\/volltexte\/2015\/2337."},{"key":"e_1_2_1_5_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR.2013.207"},{"key":"e_1_2_1_6_1","doi-asserted-by":"crossref","unstructured":"A. Babenko A. Slesarev A. Chigorin and V. Lempitsky. 2014. Neural codes for image retrieval. In Computer Vision--ECCV 2014. Springer 584--599.  A. Babenko A. Slesarev A. Chigorin and V. Lempitsky. 2014. Neural codes for image retrieval. In Computer Vision--ECCV 2014. Springer 584--599.","DOI":"10.1007\/978-3-319-10590-1_38"},{"key":"e_1_2_1_7_1","doi-asserted-by":"publisher","DOI":"10.1016\/j.is.2013.05.010"},{"key":"e_1_2_1_8_1","doi-asserted-by":"publisher","DOI":"10.1007\/11744023_32"},{"volume-title":"Pattern Recognition and Machine Learning","author":"Bishop C. M.","key":"e_1_2_1_9_1","unstructured":"C. M. Bishop . 2006. Pattern Recognition and Machine Learning . Springer . C. M. Bishop. 2006. Pattern Recognition and Machine Learning. Springer."},{"key":"e_1_2_1_10_1","unstructured":"V. Chandrasekhar J. Lin O. Mor\u00e8re H. Goh and A. Veillard. 2015. A practical guide to CNNs and fisher vectors for image instance retrieval. CoRR abs\/1508.02496 (2015). http:\/\/arxiv.org\/abs\/1508.02496  V. Chandrasekhar J. Lin O. Mor\u00e8re H. Goh and A. Veillard. 2015. A practical guide to CNNs and fisher vectors for image instance retrieval. CoRR abs\/1508.02496 (2015). http:\/\/arxiv.org\/abs\/1508.02496"},{"volume-title":"Proceedings of the 2011 Conference Record of the 45th Asilomar Conference on Signals, Systems and Computers (ASILOMAR\u201911)","author":"Chen D.","key":"e_1_2_1_11_1","unstructured":"D. Chen , S. Tsai , V. Chandrasekhar , G. Takacs , Huizhong Chen , R. Vedantham , R. Grzeszczuk , and B. Girod . 2011. Residual enhanced visual vectors for on-device image matching . In Proceedings of the 2011 Conference Record of the 45th Asilomar Conference on Signals, Systems and Computers (ASILOMAR\u201911) . 850--854. D. Chen, S. Tsai, V. Chandrasekhar, G. Takacs, Huizhong Chen, R. Vedantham, R. Grzeszczuk, and B. Girod. 2011. Residual enhanced visual vectors for on-device image matching. In Proceedings of the 2011 Conference Record of the 45th Asilomar Conference on Signals, Systems and Computers (ASILOMAR\u201911). 850--854."},{"key":"e_1_2_1_12_1","volume-title":"Proceedings of the Workshop on Statistical Learning in Computer Vision, ECCV 1, 1--22","author":"Csurka G.","year":"2004","unstructured":"G. Csurka , C. Dance , L. Fan , J. Willamowski , and C. Bray . 2004. Visual categorization with bags of keypoints . In Proceedings of the Workshop on Statistical Learning in Computer Vision, ECCV 1, 1--22 ( 2004 ), 1--2. G. Csurka, C. Dance, L. Fan, J. Willamowski, and C. Bray. 2004. Visual categorization with bags of keypoints. In Proceedings of the Workshop on Statistical Learning in Computer Vision, ECCV 1, 1--22 (2004), 1--2."},{"key":"e_1_2_1_13_1","doi-asserted-by":"publisher","DOI":"10.1145\/2502081.2502171"},{"key":"e_1_2_1_14_1","volume-title":"Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition","author":"Deng J.","year":"2009","unstructured":"J. Deng , W. Dong , R. Socher , L. J. Li , K. Li , and L. Fei-Fei . 2009. ImageNet: A large-scale hierarchical image database . In Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition , 2009 (CVPR\u201909). 248--255. J. Deng, W. Dong, R. Socher, L. J. Li, K. Li, and L. Fei-Fei. 2009. ImageNet: A large-scale hierarchical image database. In Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition, 2009 (CVPR\u201909). 248--255."},{"key":"e_1_2_1_15_1","unstructured":"J. Donahue Y. Jia O. Vinyals J. Hoffman N. Zhang E. Tzeng and T. Darrell. 2013. DeCAF: A deep convolutional activation feature for generic visual recognition. CoRR abs\/1310.1531 (2013). http:\/\/arxiv.org\/abs\/1310.1531  J. Donahue Y. Jia O. Vinyals J. Hoffman N. Zhang E. Tzeng and T. Darrell. 2013. DeCAF: A deep convolutional activation feature for generic visual recognition. CoRR abs\/1310.1531 (2013). http:\/\/arxiv.org\/abs\/1310.1531"},{"key":"e_1_2_1_16_1","doi-asserted-by":"publisher","DOI":"10.1145\/358669.358692"},{"key":"e_1_2_1_17_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR.2014.81"},{"key":"e_1_2_1_18_1","unstructured":"I. Goodfellow Y. Bengio and A. Courville. 2016. Deep Learning. Retrieved from http:\/\/www.deeplearningbook.org.  I. Goodfellow Y. Bengio and A. Courville. 2016. Deep Learning. Retrieved from http:\/\/www.deeplearningbook.org."},{"key":"e_1_2_1_19_1","doi-asserted-by":"publisher","DOI":"10.1109\/18.720541"},{"key":"e_1_2_1_20_1","unstructured":"T. Jaakkola and D. Haussler. 1998. Exploiting generative models in discriminative classifiers. In Advances in Neural Information Processing Systems 11. MIT Press 487--493.   T. Jaakkola and D. Haussler. 1998. Exploiting generative models in discriminative classifiers. In Advances in Neural Information Processing Systems 11. MIT Press 487--493."},{"key":"e_1_2_1_21_1","doi-asserted-by":"publisher","DOI":"10.1007\/s11263-009-0285-2"},{"key":"e_1_2_1_22_1","doi-asserted-by":"publisher","DOI":"10.1109\/TPAMI.2010.57"},{"volume-title":"Proceedings of the IEEE Conference on Computer Vision 8 Pattern Recognition.","author":"J\u00e9gou H.","key":"e_1_2_1_23_1","unstructured":"H. J\u00e9gou , M. Douze , C. Schmid , and P. P\u00e9rez . 2010. Aggregating local descriptors into a compact image representation . In Proceedings of the IEEE Conference on Computer Vision 8 Pattern Recognition. H. J\u00e9gou, M. Douze, C. Schmid, and P. P\u00e9rez. 2010. Aggregating local descriptors into a compact image representation. In Proceedings of the IEEE Conference on Computer Vision 8 Pattern Recognition."},{"key":"e_1_2_1_24_1","doi-asserted-by":"publisher","DOI":"10.1109\/TPAMI.2011.235"},{"key":"e_1_2_1_25_1","doi-asserted-by":"publisher","DOI":"10.1145\/2647868.2654889"},{"key":"e_1_2_1_26_1","doi-asserted-by":"publisher","DOI":"10.1145\/2390790.2390795"},{"key":"e_1_2_1_27_1","unstructured":"A. Krizhevsky I. Sutskever and G. E. Hinton. 2012. ImageNet classification with deep convolutional neural networks. In Advances in Neural Information Processing Systems 25 F. Pereira C. J. C. Burges L. Bottou and K. Q. Weinberger (Eds.). Curran Associates 1097--1105.   A. Krizhevsky I. Sutskever and G. E. Hinton. 2012. ImageNet classification with deep convolutional neural networks. In Advances in Neural Information Processing Systems 25 F. Pereira C. J. C. Burges L. Bottou and K. Q. Weinberger (Eds.). Curran Associates 1097--1105."},{"key":"e_1_2_1_28_1","doi-asserted-by":"publisher","DOI":"10.1109\/TIT.1982.1056489"},{"key":"e_1_2_1_29_1","doi-asserted-by":"publisher","DOI":"10.1023\/B:VISI.0000029664.99615.94"},{"key":"e_1_2_1_30_1","doi-asserted-by":"crossref","unstructured":"G. McLachlan and D. Peel. 2000. Finite Mixture Models. Wiley.  G. McLachlan and D. Peel. 2000. Finite Mixture Models. Wiley.","DOI":"10.1002\/0471721182"},{"key":"e_1_2_1_31_1","volume-title":"Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition","author":"Perronnin F.","year":"2007","unstructured":"F. Perronnin and C. Dance . 2007. Fisher kernels on visual vocabularies for image categorization . In Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition , 2007 (CVPR\u201907). 1--8. F. Perronnin and C. Dance. 2007. Fisher kernels on visual vocabularies for image categorization. In Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition, 2007 (CVPR\u201907). 1--8."},{"volume-title":"Proceedings of the 2010 IEEE Conference on Computer Vision and Pattern Recognition (CVPR\u201910)","author":"Perronnin F.","key":"e_1_2_1_32_1","unstructured":"F. Perronnin , Yan Liu , J. S\u00e0nchez , and H. Poirier . 2010a. Large-scale image retrieval with compressed Fisher vectors . In Proceedings of the 2010 IEEE Conference on Computer Vision and Pattern Recognition (CVPR\u201910) . 3384--3391. F. Perronnin, Yan Liu, J. S\u00e0nchez, and H. Poirier. 2010a. Large-scale image retrieval with compressed Fisher vectors. In Proceedings of the 2010 IEEE Conference on Computer Vision and Pattern Recognition (CVPR\u201910). 3384--3391."},{"key":"e_1_2_1_33_1","volume-title":"Proceedings of the Computer Vision (ECCV\u201910)","volume":"6314","author":"Perronnin F.","unstructured":"F. Perronnin , J. S\u00e0nchez , and T. Mensink . 2010b. Improving the Fisher kernel for large-scale image classification . In Proceedings of the Computer Vision (ECCV\u201910) . Lecture Notes in Computer Science , Vol. 6314 . Springer, Berlin, 143--156. F. Perronnin, J. S\u00e0nchez, and T. Mensink. 2010b. Improving the Fisher kernel for large-scale image classification. In Proceedings of the Computer Vision (ECCV\u201910). Lecture Notes in Computer Science, Vol. 6314. Springer, Berlin, 143--156."},{"volume-title":"Proceedings of the 2007 IEEE Conference on Computer Vision and Pattern Recognition (CVPR\u201907)","author":"Philbin J.","key":"e_1_2_1_34_1","unstructured":"J. Philbin , O. Chum , M. Isard , J. Sivic , and A. Zisserman . 2007. Object retrieval with large vocabularies and fast spatial matching . In Proceedings of the 2007 IEEE Conference on Computer Vision and Pattern Recognition (CVPR\u201907) . 1--8. J. Philbin, O. Chum, M. Isard, J. Sivic, and A. Zisserman. 2007. Object retrieval with large vocabularies and fast spatial matching. In Proceedings of the 2007 IEEE Conference on Computer Vision and Pattern Recognition (CVPR\u201907). 1--8."},{"volume-title":"Proceedings of the 2008 IEEE Conference on Computer Vision and Pattern Recognition (CVPR\u201908)","author":"Philbin J.","key":"e_1_2_1_35_1","unstructured":"J. Philbin , O. Chum , M. Isard , J. Sivic , and A. Zisserman . 2008. Lost in quantization: Improving particular object retrieval in large scale image databases . In Proceedings of the 2008 IEEE Conference on Computer Vision and Pattern Recognition (CVPR\u201908) . 1--8. J. Philbin, O. Chum, M. Isard, J. Sivic, and A. Zisserman. 2008. Lost in quantization: Improving particular object retrieval in large scale image databases. In Proceedings of the 2008 IEEE Conference on Computer Vision and Pattern Recognition (CVPR\u201908). 1--8."},{"key":"e_1_2_1_36_1","doi-asserted-by":"publisher","DOI":"10.1257\/jep.29.3.51"},{"key":"e_1_2_1_37_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPRW.2014.131"},{"volume-title":"Introduction to Modern Information Retrieval","author":"Salton G.","key":"e_1_2_1_38_1","unstructured":"G. Salton and M. J. McGill . 1986. Introduction to Modern Information Retrieval . McGraw-Hill , New York, NY . G. Salton and M. J. McGill. 1986. Introduction to Modern Information Retrieval. McGraw-Hill, New York, NY."},{"key":"e_1_2_1_39_1","doi-asserted-by":"publisher","DOI":"10.1007\/s11263-013-0636-x"},{"key":"e_1_2_1_40_1","unstructured":"K. Simonyan and A. Zisserman. 2014. Very deep convolutional networks for large-scale image recognition. CoRR abs\/1409.1556 (2014). http:\/\/arxiv.org\/abs\/1409.1556  K. Simonyan and A. Zisserman. 2014. Very deep convolutional networks for large-scale image recognition. CoRR abs\/1409.1556 (2014). http:\/\/arxiv.org\/abs\/1409.1556"},{"key":"e_1_2_1_41_1","volume-title":"Proceedings of the 9th IEEE International Conference on Computer Vision (ICCV\u201903)","volume":"2","author":"Sivic J.","unstructured":"J. Sivic and A. Zisserman . 2003. Video google: A text retrieval approach to object matching in videos . In Proceedings of the 9th IEEE International Conference on Computer Vision (ICCV\u201903) , Vol. 2 . IEEE Computer Society, 1470--1477. J. Sivic and A. Zisserman. 2003. Video google: A text retrieval approach to object matching in videos. In Proceedings of the 9th IEEE International Conference on Computer Vision (ICCV\u201903), Vol. 2. IEEE Computer Society, 1470--1477."},{"key":"e_1_2_1_42_1","doi-asserted-by":"publisher","DOI":"10.1145\/1873951.1874250"},{"key":"e_1_2_1_43_1","volume-title":"Research Report RR-8325.","author":"Tolias G.","year":"2013","unstructured":"G. Tolias and H. J\u00e9gou . 2013 . Local Visual Query Expansion: Exploiting an Image Collection to Refine Local Descriptors . Research Report RR-8325. Retrieved from https:\/\/hal.inria.fr\/hal-00840721. G. Tolias and H. J\u00e9gou. 2013. Local Visual Query Expansion: Exploiting an Image Collection to Refine Local Descriptors. Research Report RR-8325. Retrieved from https:\/\/hal.inria.fr\/hal-00840721."},{"key":"e_1_2_1_44_1","unstructured":"G. Tolias R. Sicre and H. J\u00e9gou. 2015. Particular object retrieval with integral max-pooling of CNN activations. arXiv preprint arXiv:1511.05879 (2015). http:\/\/arxiv.org\/abs\/1511.05879  G. Tolias R. Sicre and H. J\u00e9gou. 2015. Particular object retrieval with integral max-pooling of CNN activations. arXiv preprint arXiv:1511.05879 (2015). http:\/\/arxiv.org\/abs\/1511.05879"},{"key":"e_1_2_1_45_1","doi-asserted-by":"publisher","DOI":"10.1109\/TPAMI.2009.132"},{"volume-title":"Proceedings of the 24th British Machine Vision Conference (BMVC\u201913)","author":"Zhao W. L.","key":"e_1_2_1_46_1","unstructured":"W. L. Zhao , H. J\u00e9gou , and G. Gravier . 2013. Oriented pooling for dense and non-dense rotation-invariant features . In Proceedings of the 24th British Machine Vision Conference (BMVC\u201913) . W. L. Zhao, H. J\u00e9gou, and G. Gravier. 2013. Oriented pooling for dense and non-dense rotation-invariant features. In Proceedings of the 24th British Machine Vision Conference (BMVC\u201913)."},{"key":"e_1_2_1_47_1","volume-title":"Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition","author":"Zheng Y. T.","year":"2009","unstructured":"Y. T. Zheng , M. Zhao , Y. Song , H. Adam , U. Buddemeier , A. Bissacco , F. Brucher , T. S. Chua , and H. Neven . 2009. Tour the world: Building a web-scale landmark recognition engine . In Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition , 2009 (CVPR\u201909). 1085--1092. Y. T. Zheng, M. Zhao, Y. Song, H. Adam, U. Buddemeier, A. Bissacco, F. Brucher, T. S. Chua, and H. Neven. 2009. Tour the world: Building a web-scale landmark recognition engine. In Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition, 2009 (CVPR\u201909). 1085--1092."},{"key":"e_1_2_1_48_1","unstructured":"B. Zhou A. Lapedriza J. Xiao A. Torralba and A. Oliva. 2014. Learning deep features for scene recognition using places database. In Advances in Neural Information Processing Systems 27 Z. Ghahramani M. Welling C. Cortes N. D. Lawrence and K.Q. Weinberger (Eds.). Curran Associates 487--495.   B. Zhou A. Lapedriza J. Xiao A. Torralba and A. Oliva. 2014. Learning deep features for scene recognition using places database. In Advances in Neural Information Processing Systems 27 Z. Ghahramani M. Welling C. Cortes N. D. Lawrence and K.Q. Weinberger (Eds.). Curran Associates 487--495."}],"container-title":["Journal on Computing and Cultural Heritage"],"original-title":[],"language":"en","link":[{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/2964911","content-type":"unspecified","content-version":"vor","intended-application":"text-mining"},{"URL":"https:\/\/dl.acm.org\/doi\/pdf\/10.1145\/2964911","content-type":"unspecified","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2025,6,18]],"date-time":"2025-06-18T03:40:01Z","timestamp":1750218001000},"score":1,"resource":{"primary":{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/2964911"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2016,12,19]]},"references-count":48,"journal-issue":{"issue":"4","published-print":{"date-parts":[[2016,12,19]]}},"alternative-id":["10.1145\/2964911"],"URL":"https:\/\/doi.org\/10.1145\/2964911","relation":{},"ISSN":["1556-4673","1556-4711"],"issn-type":[{"type":"print","value":"1556-4673"},{"type":"electronic","value":"1556-4711"}],"subject":[],"published":{"date-parts":[[2016,12,19]]},"assertion":[{"value":"2016-02-01","order":0,"name":"received","label":"Received","group":{"name":"publication_history","label":"Publication History"}},{"value":"2016-06-01","order":1,"name":"accepted","label":"Accepted","group":{"name":"publication_history","label":"Publication History"}},{"value":"2016-12-19","order":2,"name":"published","label":"Published","group":{"name":"publication_history","label":"Publication History"}}]}}