{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2026,5,6]],"date-time":"2026-05-06T20:14:24Z","timestamp":1778098464322,"version":"3.51.4"},"publisher-location":"New York, NY, USA","reference-count":42,"publisher":"ACM","license":[{"start":{"date-parts":[[2020,10,12]],"date-time":"2020-10-12T00:00:00Z","timestamp":1602460800000},"content-version":"vor","delay-in-days":0,"URL":"https:\/\/www.acm.org\/publications\/policies\/copyright_policy#Background"}],"content-domain":{"domain":["dl.acm.org"],"crossmark-restriction":true},"short-container-title":[],"published-print":{"date-parts":[[2020,10,12]]},"DOI":"10.1145\/3423323.3423409","type":"proceedings-article","created":{"date-parts":[[2020,10,12]],"date-time":"2020-10-12T09:27:36Z","timestamp":1602494856000},"page":"5-13","update-policy":"https:\/\/doi.org\/10.1145\/crossmark-policy","source":"Crossref","is-referenced-by-count":3,"title":["Semantics Preserving Hierarchy based Retrieval of Indian heritage monuments"],"prefix":"10.1145","author":[{"given":"Ronak","family":"Gupta","sequence":"first","affiliation":[{"name":"Indian Institute of Technology Delhi, Delhi, India"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Prerana","family":"Mukherjee","sequence":"additional","affiliation":[{"name":"Jawaharlal Nehru University, Delhi, India"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Brejesh","family":"Lall","sequence":"additional","affiliation":[{"name":"Indian Institute of Technology Delhi, Delhi, India"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Varshul","family":"Gupta","sequence":"additional","affiliation":[{"name":"Jaypee Institute of Information Technology, Delhi, India"}],"role":[{"role":"author","vocabulary":"crossref"}]}],"member":"320","published-online":{"date-parts":[[2020,10,12]]},"reference":[{"key":"e_1_3_2_1_1_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR.2016.385"},{"key":"e_1_3_2_1_2_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR.2012.6248018"},{"key":"e_1_3_2_1_3_1","volume-title":"Proceedings of the IEEE international conference on computer vision. 1269--1277","author":"Babenko Artem","year":"2015","unstructured":"Artem Babenko and Victor Lempitsky . 2015 . Aggregating local deep features for image retrieval . In Proceedings of the IEEE international conference on computer vision. 1269--1277 . Artem Babenko and Victor Lempitsky. 2015. Aggregating local deep features for image retrieval. In Proceedings of the IEEE international conference on computer vision. 1269--1277."},{"key":"e_1_3_2_1_4_1","doi-asserted-by":"publisher","DOI":"10.1109\/WACV.2019.00073"},{"key":"e_1_3_2_1_5_1","doi-asserted-by":"publisher","DOI":"10.1016\/j.cviu.2014.11.006"},{"key":"e_1_3_2_1_6_1","volume-title":"Multimedia ontology: representation and applications","author":"Chaudhury Santanu","unstructured":"Santanu Chaudhury , Anupama Mallik , and Hiranmay Ghosh . 2015. Multimedia ontology: representation and applications . CRC Press . Santanu Chaudhury, Anupama Mallik, and Hiranmay Ghosh. 2015. Multimedia ontology: representation and applications. CRC Press."},{"key":"e_1_3_2_1_7_1","doi-asserted-by":"publisher","DOI":"10.1109\/ICCVW.2017.347"},{"key":"e_1_3_2_1_8_1","volume-title":"ECCV","volume":"1","author":"Csurka Gabriella","year":"2004","unstructured":"Gabriella Csurka , Christopher Dance , Lixin Fan , Jutta Willamowski , and C\u00e9dric Bray . 2004 . Visual categorization with bags of keypoints. InWorkshop on statistical learning in computer vision , ECCV , Vol. 1 . Prague, 1--2. Gabriella Csurka, Christopher Dance, Lixin Fan, Jutta Willamowski, and C\u00e9dric Bray. 2004. Visual categorization with bags of keypoints. InWorkshop on statistical learning in computer vision, ECCV, Vol. 1. Prague, 1--2."},{"key":"e_1_3_2_1_9_1","doi-asserted-by":"publisher","DOI":"10.1145\/2502081.2502171"},{"key":"e_1_3_2_1_10_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR.2011.5995516"},{"key":"e_1_3_2_1_11_1","doi-asserted-by":"publisher","DOI":"10.1145\/2830541"},{"key":"e_1_3_2_1_12_1","volume-title":"Proceedings of the Eighth Indian Conference on Computer Vision, Graphics and Image Processing. 1--8.","author":"Goel Abhinav","year":"2012","unstructured":"Abhinav Goel , Mayank Juneja , and CV Jawahar . 2012 . Are buildings only instances? Exploration in architectural style categories . In Proceedings of the Eighth Indian Conference on Computer Vision, Graphics and Image Processing. 1--8. Abhinav Goel, Mayank Juneja, and CV Jawahar. 2012. Are buildings only instances? Exploration in architectural style categories. In Proceedings of the Eighth Indian Conference on Computer Vision, Graphics and Image Processing. 1--8."},{"key":"e_1_3_2_1_13_1","doi-asserted-by":"publisher","DOI":"10.1007\/978-3-319-46466-4_15"},{"key":"e_1_3_2_1_14_1","doi-asserted-by":"publisher","DOI":"10.1007\/s11263-017-1016-8"},{"key":"e_1_3_2_1_15_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR.2016.90"},{"key":"e_1_3_2_1_16_1","volume-title":"Class-Weighted Convolutional Features for Visual Instance Search. In 28th British Machine Vision Conference (BMVC).","author":"Jimenez Albert","year":"2017","unstructured":"Albert Jimenez , Jose M. Alvarez , and Xavier Giro-i Nieto . 2017 . Class-Weighted Convolutional Features for Visual Instance Search. In 28th British Machine Vision Conference (BMVC). Albert Jimenez, Jose M. Alvarez, and Xavier Giro-i Nieto. 2017. Class-Weighted Convolutional Features for Visual Instance Search. In 28th British Machine Vision Conference (BMVC)."},{"key":"e_1_3_2_1_17_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR.2015.7298990"},{"key":"e_1_3_2_1_18_1","unstructured":"Jaeyoon Kim and Sung-Eui Yoon. 2018. Regional Attention Based Deep Feature for Image Retrieval.. In BMVC. 209.  Jaeyoon Kim and Sung-Eui Yoon. 2018. Regional Attention Based Deep Feature for Image Retrieval.. In BMVC. 209."},{"key":"e_1_3_2_1_19_1","unstructured":"Alex Krizhevsky Ilya Sutskever and Geoffrey E Hinton. 2012. Imagenet classification with deep convolutional neural networks. In Advances in neural information processing systems. 1097--1105.  Alex Krizhevsky Ilya Sutskever and Geoffrey E Hinton. 2012. Imagenet classification with deep convolutional neural networks. In Advances in neural information processing systems. 1097--1105."},{"key":"e_1_3_2_1_20_1","volume-title":"Improving Landmark Recognition using Saliency detection and Feature classification. arXiv preprint arXiv:1811.12748","author":"Kumar Akash","year":"2018","unstructured":"Akash Kumar , Sagnik Bhowmick , N Jayanthi , and S Indu . 2018. Improving Landmark Recognition using Saliency detection and Feature classification. arXiv preprint arXiv:1811.12748 ( 2018 ). Akash Kumar, Sagnik Bhowmick, N Jayanthi, and S Indu. 2018. Improving Landmark Recognition using Saliency detection and Feature classification. arXiv preprint arXiv:1811.12748 (2018)."},{"key":"e_1_3_2_1_21_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPRW.2015.7301269"},{"key":"e_1_3_2_1_22_1","volume-title":"Sgdr: Stochastic gradient descent with warm restarts. arXiv preprint arXiv:1608.03983","author":"Loshchilov Ilya","year":"2016","unstructured":"Ilya Loshchilov and Frank Hutter . 2016 . Sgdr: Stochastic gradient descent with warm restarts. arXiv preprint arXiv:1608.03983 (2016). Ilya Loshchilov and Frank Hutter. 2016. Sgdr: Stochastic gradient descent with warm restarts. arXiv preprint arXiv:1608.03983 (2016)."},{"key":"e_1_3_2_1_23_1","doi-asserted-by":"publisher","DOI":"10.1109\/TIP.2016.2627801"},{"key":"e_1_3_2_1_24_1","volume-title":"Proceedings of the 1stWorkshop on Structuring and Understanding of Multimedia heritAge Contents. 55--59","author":"Obeso Abraham Montoya","year":"2019","unstructured":"Abraham Montoya Obeso , Jenny Benois-Pineau , Mireya Sara\u00ed Garc\u00eda V\u00e1zquez , and Alejandro \u00c1lvaro Ram\u00edrez Acosta . 2019 . Organizing cultural heritage with deep features . In Proceedings of the 1stWorkshop on Structuring and Understanding of Multimedia heritAge Contents. 55--59 . Abraham Montoya Obeso, Jenny Benois-Pineau, Mireya Sara\u00ed Garc\u00eda V\u00e1zquez, and Alejandro \u00c1lvaro Ram\u00edrez Acosta. 2019. Organizing cultural heritage with deep features. In Proceedings of the 1stWorkshop on Structuring and Understanding of Multimedia heritAge Contents. 55--59."},{"key":"e_1_3_2_1_25_1","doi-asserted-by":"publisher","DOI":"10.1109\/ICCV.2017.374"},{"key":"e_1_3_2_1_26_1","doi-asserted-by":"publisher","DOI":"10.1109\/IPTA.2019.8936125"},{"key":"e_1_3_2_1_27_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR.2007.383172"},{"key":"e_1_3_2_1_28_1","volume-title":"Jayanta Mukherjee, and Shamik Sural.","author":"Podder Dipannita","year":"2018","unstructured":"Dipannita Podder , Jit Mukherjee , Shashaank Mattur Aswatha , Jayanta Mukherjee, and Shamik Sural. 2018 . Ontology-driven content-based retrieval of heritage images. In Heritage Preservation. Springer , 143--160. Dipannita Podder, Jit Mukherjee, Shashaank Mattur Aswatha, Jayanta Mukherjee, and Shamik Sural. 2018. Ontology-driven content-based retrieval of heritage images. In Heritage Preservation. Springer, 143--160."},{"key":"e_1_3_2_1_29_1","volume-title":"Fine-tuning CNN image retrieval with no human annotation","author":"Radenovic Filip","year":"2018","unstructured":"Filip Radenovic , Giorgos Tolias , and Ondrej Chum . 2018. Fine-tuning CNN image retrieval with no human annotation . IEEE transactions on pattern analysis and machine intelligence ( 2018 ). Filip Radenovic, Giorgos Tolias, and Ondrej Chum. 2018. Fine-tuning CNN image retrieval with no human annotation. IEEE transactions on pattern analysis and machine intelligence (2018)."},{"key":"e_1_3_2_1_30_1","doi-asserted-by":"publisher","DOI":"10.1109\/ICCV.2011.6126544"},{"key":"e_1_3_2_1_31_1","doi-asserted-by":"crossref","unstructured":"Olga Russakovsky Jia Deng Hao Su Jonathan Krause Sanjeev Satheesh Sean Ma Zhiheng Huang Andrej Karpathy Aditya Khosla Michael Bernstein etal 2015. Imagenet large scale visual recognition challenge. International journal of computer vision 115 3 (2015) 211--252.  Olga Russakovsky Jia Deng Hao Su Jonathan Krause Sanjeev Satheesh Sean Ma Zhiheng Huang Andrej Karpathy Aditya Khosla Michael Bernstein et al. 2015. Imagenet large scale visual recognition challenge. International journal of computer vision 115 3 (2015) 211--252.","DOI":"10.1007\/s11263-015-0816-y"},{"key":"e_1_3_2_1_32_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPRW.2016.56"},{"key":"e_1_3_2_1_33_1","volume-title":"International journal of computer vision 105, 3","author":"S\u00e1nchez Jorge","year":"2013","unstructured":"Jorge S\u00e1nchez , Florent Perronnin , Thomas Mensink , and Jakob Verbeek . 2013. Image classification with the fisher vector: Theory and practice . International journal of computer vision 105, 3 ( 2013 ), 222--245. Jorge S\u00e1nchez, Florent Perronnin, Thomas Mensink, and Jakob Verbeek. 2013. Image classification with the fisher vector: Theory and practice. International journal of computer vision 105, 3 (2013), 222--245."},{"key":"e_1_3_2_1_34_1","doi-asserted-by":"publisher","DOI":"10.1109\/SITIS.2017.29"},{"key":"e_1_3_2_1_35_1","volume-title":"Video Google: A text retrieval approach to object matching in videos. In null","author":"Sivic Josef","year":"2003","unstructured":"Josef Sivic and Andrew Zisserman . 2003 . Video Google: A text retrieval approach to object matching in videos. In null . IEEE , 1470. Josef Sivic and Andrew Zisserman. 2003. Video Google: A text retrieval approach to object matching in videos. In null. IEEE, 1470."},{"key":"e_1_3_2_1_36_1","volume-title":"Thirty-Second AAAI Conference on Artificial Intelligence.","author":"Song Jingkuan","year":"2018","unstructured":"Jingkuan Song , Tao He , Lianli Gao , Xing Xu , Alan Hanjalic , and Heng Tao Shen . 2018 . Binary generative adversarial networks for image retrieval . In Thirty-Second AAAI Conference on Artificial Intelligence. Jingkuan Song, Tao He, Lianli Gao, Xing Xu, Alan Hanjalic, and Heng Tao Shen. 2018. Binary generative adversarial networks for image retrieval. In Thirty-Second AAAI Conference on Artificial Intelligence."},{"key":"e_1_3_2_1_37_1","volume-title":"Label embedding network: Learning label representation for soft training of deep networks. arXiv preprint arXiv:1710.10393","author":"Sun Xu","year":"2017","unstructured":"Xu Sun , BingzhenWei, Xuancheng Ren , and Shuming Ma. 2017. Label embedding network: Learning label representation for soft training of deep networks. arXiv preprint arXiv:1710.10393 ( 2017 ). Xu Sun, BingzhenWei, Xuancheng Ren, and Shuming Ma. 2017. Label embedding network: Learning label representation for soft training of deep networks. arXiv preprint arXiv:1710.10393 (2017)."},{"key":"e_1_3_2_1_38_1","volume-title":"Detectto- Retrieve: Efficient Regional Aggregation for Image Search. arXiv preprint arXiv:1812.01584","author":"Teichmann Marvin","year":"2018","unstructured":"Marvin Teichmann , Andre Araujo , Menglong Zhu , and Jack Sim . 2018 . Detectto- Retrieve: Efficient Regional Aggregation for Image Search. arXiv preprint arXiv:1812.01584 (2018). Marvin Teichmann, Andre Araujo, Menglong Zhu, and Jack Sim. 2018. Detectto- Retrieve: Efficient Regional Aggregation for Image Search. arXiv preprint arXiv:1812.01584 (2018)."},{"key":"e_1_3_2_1_39_1","volume-title":"Particular object retrieval with integral max-pooling of CNN activations. arXiv preprint arXiv:1511.05879","author":"Tolias Giorgos","year":"2015","unstructured":"Giorgos Tolias , Ronan Sicre , and Herv\u00e9 J\u00e9gou . 2015. Particular object retrieval with integral max-pooling of CNN activations. arXiv preprint arXiv:1511.05879 ( 2015 ). Giorgos Tolias, Ronan Sicre, and Herv\u00e9 J\u00e9gou. 2015. Particular object retrieval with integral max-pooling of CNN activations. arXiv preprint arXiv:1511.05879 (2015)."},{"key":"e_1_3_2_1_40_1","doi-asserted-by":"publisher","DOI":"10.1016\/j.neucom.2017.05.099"},{"key":"e_1_3_2_1_41_1","unstructured":"Bin Zhao Fei Li and Eric P Xing. 2011. Large-scale category structure aware image categorization. In Advances in Neural Information Processing Systems. 1251--1259.  Bin Zhao Fei Li and Eric P Xing. 2011. Large-scale category structure aware image categorization. In Advances in Neural Information Processing Systems. 1251--1259."},{"key":"e_1_3_2_1_42_1","volume-title":"Proceedings of the IEEE conference on computer vision and pattern recognition. 1556--1564","author":"Zhao Fang","year":"2015","unstructured":"Fang Zhao , Yongzhen Huang , Liang Wang , and Tieniu Tan . 2015 . Deep semantic ranking based hashing for multi-label image retrieval . In Proceedings of the IEEE conference on computer vision and pattern recognition. 1556--1564 . Fang Zhao, Yongzhen Huang, Liang Wang, and Tieniu Tan. 2015. Deep semantic ranking based hashing for multi-label image retrieval. In Proceedings of the IEEE conference on computer vision and pattern recognition. 1556--1564."}],"event":{"name":"MM '20: The 28th ACM International Conference on Multimedia","location":"Seattle WA USA","acronym":"MM '20","sponsor":["SIGMM ACM Special Interest Group on Multimedia"]},"container-title":["Proceedings of the 2nd Workshop on Structuring and Understanding of Multimedia heritAge Contents"],"original-title":[],"link":[{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/3423323.3423409","content-type":"unspecified","content-version":"vor","intended-application":"text-mining"},{"URL":"https:\/\/dl.acm.org\/doi\/pdf\/10.1145\/3423323.3423409","content-type":"unspecified","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2025,6,17]],"date-time":"2025-06-17T21:24:57Z","timestamp":1750195497000},"score":1,"resource":{"primary":{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/3423323.3423409"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2020,10,12]]},"references-count":42,"alternative-id":["10.1145\/3423323.3423409","10.1145\/3423323"],"URL":"https:\/\/doi.org\/10.1145\/3423323.3423409","relation":{},"subject":[],"published":{"date-parts":[[2020,10,12]]},"assertion":[{"value":"2020-10-12","order":2,"name":"published","label":"Published","group":{"name":"publication_history","label":"Publication History"}}]}}