{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2026,6,10]],"date-time":"2026-06-10T16:35:43Z","timestamp":1781109343386,"version":"3.54.1"},"reference-count":27,"publisher":"IGI Global Scientific Publishing","issue":"2","content-domain":{"domain":[],"crossmark-restriction":false},"short-container-title":[],"published-print":{"date-parts":[[2019,4]]},"abstract":"<jats:p>The authors present a robust and extendable localization system for monocular images. To have both robustness toward noise factors and extendibility to unfamiliar scenes simultaneously, our system combines traditional content-based image retrieval structure with CNN feature extraction model to localize monocular images. The core model of the system is a deep CNN feature extraction model. The feature extraction model can map an image to a d-dimension space where image pairs in the real word have smaller Euclidean distances. The feature extraction model is achieved using a deep Convnet modified from GoogLeNet. A special way to train the feature extraction model is proposed in the article using localization results from Cambridge Landmarks dataset. Through experiments, it is shown that the system is robust to noise factors supported by high level CNN features. Furthermore, the authors show that the system has a powerful extendibility to other unfamiliar scenes supported by a feature extract model's generic property and structure.<\/jats:p>","DOI":"10.4018\/ijssci.2019040103","type":"journal-article","created":{"date-parts":[[2019,7,10]],"date-time":"2019-07-10T11:32:51Z","timestamp":1562758371000},"page":"38-50","source":"Crossref","is-referenced-by-count":6,"title":["A Novel Convolutional Neural Network Based Localization System for Monocular Images"],"prefix":"10.4018","volume":"11","author":[{"given":"Chen","family":"Sun","sequence":"first","affiliation":[{"name":"Tsinghua University, Beijing, China"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Chunping","family":"Li","sequence":"additional","affiliation":[{"name":"Tsinghua University, Beijing, China"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"ORCID":"https:\/\/orcid.org\/0000-0002-5804-8926","authenticated-orcid":true,"given":"Yan","family":"Zhu","sequence":"additional","affiliation":[{"name":"Southwest Jiaotong University, Chengdu, China"}],"role":[{"vocabulary":"crossref","role":"author"}]}],"member":"2432","reference":[{"key":"IJSSCI.2019040103-0","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR.2013.104"},{"key":"IJSSCI.2019040103-1","doi-asserted-by":"publisher","DOI":"10.1109\/ICRA.2017.7989366"},{"key":"IJSSCI.2019040103-2","doi-asserted-by":"publisher","DOI":"10.1177\/0278364908090961"},{"issue":"1","key":"IJSSCI.2019040103-3","first-page":"815","article-title":"Decaf: A deep convolutional activation feature for generic visual recognition.","volume":"50","author":"J.Donahue","year":"2013","journal-title":"Computer Science"},{"key":"IJSSCI.2019040103-4","doi-asserted-by":"crossref","unstructured":"Engel, J., Schps, T., & Cremers, D. (2014). LSD-SLAM: Large-Scale Direct Monocular SLAM. In Proceedings of theEuropean Conference on Computer Vision (pp. 834-849). Academic Press.","DOI":"10.1007\/978-3-319-10605-2_54"},{"key":"IJSSCI.2019040103-5","unstructured":"Gionis, A., Indyk, P., & Motwani, R. (2000). Similarity search in high dimensions via hashing. In Proceedings of theInternational Conference on Very Large Data Bases (pp. 518\u2013529). Academic Press."},{"key":"IJSSCI.2019040103-6","doi-asserted-by":"publisher","DOI":"10.1177\/0278364911430419"},{"key":"IJSSCI.2019040103-7","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR.2017.694"},{"key":"IJSSCI.2019040103-8","unstructured":"Kendall, A., Gal, Y., & Cipolla, R. (2018). Multi-Task Learning Using Uncertainty to Weigh Losses for Scene Geometry and Semantics. In Proc. CVPR (Vol. 1). Academic Press."},{"key":"IJSSCI.2019040103-9","first-page":"125","article-title":"Posenet: A convolutional network for real-time 6-dof camera relocalization.","volume":"31","author":"A.Kendall","year":"2016","journal-title":"Education for Information"},{"key":"IJSSCI.2019040103-10","doi-asserted-by":"publisher","DOI":"10.1109\/ISMAR.2007.4538852"},{"key":"IJSSCI.2019040103-11","doi-asserted-by":"publisher","DOI":"10.1007\/978-3-642-33718-5_2"},{"key":"IJSSCI.2019040103-12","doi-asserted-by":"publisher","DOI":"10.1023\/B:VISI.0000029664.99615.94"},{"key":"IJSSCI.2019040103-13","unstructured":"Michal, W.F. (1995). Orientation histograms for hand gesture recognition. Mitsubishi Electric Research Labs."},{"key":"IJSSCI.2019040103-14","doi-asserted-by":"publisher","DOI":"10.1109\/ICCV.2011.6126513"},{"key":"IJSSCI.2019040103-15","doi-asserted-by":"publisher","DOI":"10.1109\/ICPR.1994.576366"},{"key":"IJSSCI.2019040103-16","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR.2014.222"},{"key":"IJSSCI.2019040103-17","doi-asserted-by":"publisher","DOI":"10.1109\/CVPRW.2014.131"},{"key":"IJSSCI.2019040103-18","doi-asserted-by":"publisher","DOI":"10.1109\/TPAMI.2016.2611662"},{"key":"IJSSCI.2019040103-19","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR.2013.377"},{"key":"IJSSCI.2019040103-20","doi-asserted-by":"publisher","DOI":"10.1109\/IROS.2015.7353986"},{"key":"IJSSCI.2019040103-21","doi-asserted-by":"crossref","unstructured":"Szegedy, C., Liu, W., Jia, Y., & Sermanet, P. 2015. Going deeper with convolutions. In Proceedings of theIEEE Conference on Computer Vision and Pattern Recognition. IEEE.","DOI":"10.1109\/CVPR.2015.7298594"},{"key":"IJSSCI.2019040103-22","doi-asserted-by":"publisher","DOI":"10.1109\/ICCV.2017.75"},{"key":"IJSSCI.2019040103-23","doi-asserted-by":"publisher","DOI":"10.1109\/TSMCB.2005.859085"},{"key":"IJSSCI.2019040103-24","doi-asserted-by":"crossref","unstructured":"Weyand, T., Kostrikov, I., & Philbin, J. 2016. Planet-photo geolocation with convolutional neural networks. In Proceedings of theEuropean Conference on Computer Vision (pp. 37\u201355). Academic Press.","DOI":"10.1007\/978-3-319-46484-8_3"},{"key":"IJSSCI.2019040103-25","unstructured":"Wu, F., Pang, Y., Zhang, L., Li, Z., Cai, R., & Hao, Q. 2012. 3D visual phrases for landmark recognition. In Proceedings of theIEEE Conference on Computer Vision and Pattern Recognition (pp. 3594-3601). IEEE."},{"key":"IJSSCI.2019040103-26","doi-asserted-by":"publisher","DOI":"10.1145\/1655925.1656049"}],"container-title":["International Journal of Software Science and Computational Intelligence"],"original-title":[],"language":"en","link":[{"URL":"https:\/\/www.igi-global.com\/viewtitle.aspx?TitleId=233522","content-type":"unspecified","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2022,5,6]],"date-time":"2022-05-06T08:06:43Z","timestamp":1651824403000},"score":1,"resource":{"primary":{"URL":"http:\/\/services.igi-global.com\/resolvedoi\/resolve.aspx?doi=10.4018\/IJSSCI.2019040103"}},"subtitle":[""],"short-title":[],"issued":{"date-parts":[[2019,4]]},"references-count":27,"journal-issue":{"issue":"2"},"URL":"https:\/\/doi.org\/10.4018\/ijssci.2019040103","relation":{},"ISSN":["1942-9045","1942-9037"],"issn-type":[{"value":"1942-9045","type":"print"},{"value":"1942-9037","type":"electronic"}],"subject":[],"published":{"date-parts":[[2019,4]]}}}