{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2026,7,20]],"date-time":"2026-07-20T17:03:22Z","timestamp":1784567002530,"version":"3.55.0"},"reference-count":151,"publisher":"Springer Science and Business Media LLC","issue":"6","license":[{"start":{"date-parts":[[2026,5,9]],"date-time":"2026-05-09T00:00:00Z","timestamp":1778284800000},"content-version":"tdm","delay-in-days":0,"URL":"https:\/\/creativecommons.org\/licenses\/by\/4.0"},{"start":{"date-parts":[[2026,5,9]],"date-time":"2026-05-09T00:00:00Z","timestamp":1778284800000},"content-version":"vor","delay-in-days":0,"URL":"https:\/\/creativecommons.org\/licenses\/by\/4.0"}],"funder":[{"DOI":"10.13039\/501100001824","name":"Grantov\u00e1 Agentura \u010cesk\u00e9 Republiky","doi-asserted-by":"publisher","award":["22-23183M"],"award-info":[{"award-number":["22-23183M"]}],"id":[{"id":"10.13039\/501100001824","id-type":"DOI","asserted-by":"publisher"}]},{"DOI":"10.13039\/501100001824","name":"Grantov\u00e1 Agentura \u010cesk\u00e9 Republiky","doi-asserted-by":"publisher","award":["22-23183M"],"award-info":[{"award-number":["22-23183M"]}],"id":[{"id":"10.13039\/501100001824","id-type":"DOI","asserted-by":"publisher"}]},{"DOI":"10.13039\/501100001824","name":"Grantov\u00e1 Agentura \u010cesk\u00e9 Republiky","doi-asserted-by":"publisher","award":["23-07973X"],"award-info":[{"award-number":["23-07973X"]}],"id":[{"id":"10.13039\/501100001824","id-type":"DOI","asserted-by":"publisher"}]},{"DOI":"10.13039\/501100001824","name":"Grantov\u00e1 Agentura \u010cesk\u00e9 Republiky","doi-asserted-by":"publisher","award":["23-07973X"],"award-info":[{"award-number":["23-07973X"]}],"id":[{"id":"10.13039\/501100001824","id-type":"DOI","asserted-by":"publisher"}]},{"DOI":"10.13039\/100019180","name":"HORIZON EUROPE European Research Council","doi-asserted-by":"publisher","award":["101043189"],"award-info":[{"award-number":["101043189"]}],"id":[{"id":"10.13039\/100019180","id-type":"DOI","asserted-by":"publisher"}]},{"DOI":"10.13039\/100007655","name":"\u010cesk\u00e9 Vysok\u00e9 U\u010den\u00ed Technick\u00e9 v Praze","doi-asserted-by":"publisher","award":["SGS23\/121\/OHK3\/2T\/13"],"award-info":[{"award-number":["SGS23\/121\/OHK3\/2T\/13"]}],"id":[{"id":"10.13039\/100007655","id-type":"DOI","asserted-by":"publisher"}]}],"content-domain":{"domain":["link.springer.com"],"crossmark-restriction":false},"short-container-title":["Int J Comput Vis"],"published-print":{"date-parts":[[2026,6]]},"abstract":"<jats:title>Abstract<\/jats:title>\n                  <jats:p>Visual localization algorithms, i.e., methods that estimate the camera pose of a query image in a known scene, are core components of many applications, including self-driving cars and augmented \/ mixed reality systems. State-of-the-art visual localization algorithms are structure-based, i.e., they store a 3D model of the scene and use 2D-3D correspondences between the query image and 3D points in the model for camera pose estimation. While such approaches are highly accurate, they are also rather inflexible when it comes to adjusting the underlying 3D model after changes in the scene. Structureless localization approaches represent the scene as a database of images with known poses and thus offer a much more flexible representation that can be easily updated by adding or removing images. Although there is a large amount of literature on structure-based approaches, there is significantly less work on structureless methods. Hence, this paper is dedicated to providing the, to the best of our knowledge, first comprehensive discussion and comparison of structureless methods. Extensive experiments show that approaches that use a higher degree of classical geometric reasoning generally achieve higher pose accuracy. In particular, approaches based on classical absolute or semi-generalized relative pose estimation outperform very recent methods based on pose regression by a wide margin. Compared with state-of-the-art structure-based approaches, the flexibility of structureless methods comes at the cost of (slightly) lower pose accuracy, indicating an interesting direction for future work.<\/jats:p>","DOI":"10.1007\/s11263-026-02780-9","type":"journal-article","created":{"date-parts":[[2026,5,9]],"date-time":"2026-05-09T05:57:58Z","timestamp":1778306278000},"update-policy":"https:\/\/doi.org\/10.1007\/springer_crossmark_policy","source":"Crossref","is-referenced-by-count":0,"title":["A Guide to Structureless Visual Localization"],"prefix":"10.1007","volume":"134","author":[{"ORCID":"https:\/\/orcid.org\/0000-0003-0601-7682","authenticated-orcid":false,"given":"Vojtech","family":"Panek","sequence":"first","affiliation":[],"role":[{"vocabulary":"crossref","role":"author"}]},{"ORCID":"https:\/\/orcid.org\/0000-0002-2434-2393","authenticated-orcid":false,"given":"Qunjie","family":"Zhou","sequence":"additional","affiliation":[],"role":[{"vocabulary":"crossref","role":"author"}]},{"ORCID":"https:\/\/orcid.org\/0000-0002-7448-6686","authenticated-orcid":false,"given":"Yaqing","family":"Ding","sequence":"additional","affiliation":[],"role":[{"vocabulary":"crossref","role":"author"}]},{"ORCID":"https:\/\/orcid.org\/0000-0001-7008-1756","authenticated-orcid":false,"given":"S\u00e9rgio","family":"Agostinho","sequence":"additional","affiliation":[],"role":[{"vocabulary":"crossref","role":"author"}]},{"ORCID":"https:\/\/orcid.org\/0000-0002-1916-8829","authenticated-orcid":false,"given":"Zuzana","family":"Kukelova","sequence":"additional","affiliation":[],"role":[{"vocabulary":"crossref","role":"author"}]},{"ORCID":"https:\/\/orcid.org\/0000-0001-9760-4553","authenticated-orcid":false,"given":"Torsten","family":"Sattler","sequence":"additional","affiliation":[],"role":[{"vocabulary":"crossref","role":"author"}]},{"ORCID":"https:\/\/orcid.org\/0000-0001-8709-1133","authenticated-orcid":false,"given":"Laura","family":"Leal-Taix\u00e9","sequence":"additional","affiliation":[],"role":[{"vocabulary":"crossref","role":"author"}]}],"member":"297","published-online":{"date-parts":[[2026,5,9]]},"reference":[{"key":"2780_CR1","doi-asserted-by":"crossref","unstructured":"Adam, A., Sattler, T., Karantzalos, K., & Pajdla, T. (2022). Objects can move: 3d change detection by geometric transformation consistency. In: European Conference on Computer Vision (ECCV), pp. 108\u2013124. Springer.","DOI":"10.1007\/978-3-031-19827-4_7"},{"key":"2780_CR2","doi-asserted-by":"crossref","unstructured":"Arandjelovic, R., & Zisserman, A. (2013). All about VLAD. In: Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition, pp. 1578\u20131585.","DOI":"10.1109\/CVPR.2013.207"},{"key":"2780_CR3","doi-asserted-by":"crossref","unstructured":"Arandjelovi\u0107, R., & Zisserman, A. (2014). Visual vocabulary with a semantic twist. In: Asian Conference on Computer Vision (ACCV).","DOI":"10.1007\/978-3-319-16865-4_12"},{"key":"2780_CR4","first-page":"5297","volume":"2016","author":"R Arandjelovi\u0107","year":"2015","unstructured":"Arandjelovi\u0107, R., Gron\u00e1t, P., Torii, A., Pajdla, T., & Sivic, J. (2015). NetVLAD: CNN Architecture for Weakly Supervised Place Recognition. IEEE Conference on Computer Vision and Pattern Recognition (CVPR), 2016, 5297\u20135307.","journal-title":"IEEE Conference on Computer Vision and Pattern Recognition (CVPR)"},{"key":"2780_CR5","doi-asserted-by":"crossref","unstructured":"Ardeshir, S., Zamir, A.R., Torroella, A., & Shah, M. (2014). GIS-Assisted Object Detection and Geospatial Localization. In: ECCV.","DOI":"10.1007\/978-3-319-10599-4_39"},{"key":"2780_CR6","doi-asserted-by":"crossref","unstructured":"Arnold, E., Wynn, J., Vicente, S., Garcia-Hernando, G., Monszpart, A., Prisacariu, V., Turmukhambetov, D., & Brachmann, E. (2022). Map-free Visual Relocalization: Metric Pose Relative to a Single Image. In: European Conference on Computer Vision (ECCV), pp. 690\u2013708. Springer.","DOI":"10.1007\/978-3-031-19769-7_40"},{"key":"2780_CR7","doi-asserted-by":"crossref","unstructured":"Arth, C., Wagner, D., Klopschitz, M., Irschara, A., & Schmalstieg, D. (2009). Wide area localization on mobile phones. In: ISMAR.","DOI":"10.1109\/ISMAR.2009.5336494"},{"key":"2780_CR8","doi-asserted-by":"crossref","unstructured":"Baatz, G., K\u00f6ser, K., Chen, D., Grzeszczuk, R., & Pollefeys, M. (2010). Handling Urban Location Recognition as a 2D Homothetic Problem. In: ECCV.","DOI":"10.1007\/978-3-642-15567-3_20"},{"key":"2780_CR9","doi-asserted-by":"publisher","first-page":"315","DOI":"10.1007\/s11263-011-0458-7","volume":"96","author":"G Baatz","year":"2011","unstructured":"Baatz, G., K\u00f6ser, K., Chen, D. M., Grzeszczuk, R., & Pollefeys, M. (2011). Leveraging 3D City Models for Rotation Invariant Place-of-Interest Recognition. IJCV, 96, 315\u2013334.","journal-title":"IJCV"},{"key":"2780_CR10","unstructured":"Badino, H., Huber, D., & Kanade, T. (2011). The CMU Visual Localization Data Set. http:\/\/3dvis.ri.cmu.edu\/data-sets\/localization."},{"key":"2780_CR11","doi-asserted-by":"crossref","unstructured":"Balntas, V., Li, S., & Prisacariu, V. (2018). RelocNet: Continuous Metric Learning Relocalisation using Neural Nets. In: The European Conference on Computer Vision (ECCV).","DOI":"10.1007\/978-3-030-01264-9_46"},{"key":"2780_CR12","first-page":"1301","volume":"2020","author":"D Bar\u00e1th","year":"2019","unstructured":"Bar\u00e1th, D., Noskova, J., Ivashechkin, M., & Matas, J. (2019). MAGSAC++, a Fast, Reliable and Accurate Robust Estimator. IEEE\/CVF Conference on Computer Vision and Pattern Recognition (CVPR), 2020, 1301\u20131309.","journal-title":"IEEE\/CVF Conference on Computer Vision and Pattern Recognition (CVPR)"},{"key":"2780_CR13","doi-asserted-by":"crossref","unstructured":"Berton, G., Masone, C., & Caputo, B. (2022). Rethinking Visual Geo-Localization for Large-Scale Applications. In: IEEE\/CVF Conference on Computer Vision and Pattern Recognition (CVPR).","DOI":"10.1109\/CVPR52688.2022.00483"},{"key":"2780_CR14","doi-asserted-by":"crossref","unstructured":"Berton, G., Trivigno, G., Caputo, B., & Masone, C. (2023). EigenPlaces: Training Viewpoint Robust Models for Visual Place Recognition. In: Proceedings of the IEEE\/CVF International Conference on Computer Vision (ICCV), pp. 11080\u201311090.","DOI":"10.1109\/ICCV51070.2023.01017"},{"key":"2780_CR15","first-page":"5916","volume":"2021","author":"S Bhayani","year":"2021","unstructured":"Bhayani, S., Sattler, T., Bar\u00e1th, D., Beliansky, P., Heikkila, J., & Kukelova, Z. (2021). Calibrated and Partially Calibrated Semi-Generalized Homographies. IEEE\/CVF International Conference on Computer Vision (ICCV), 2021, 5916\u20135925.","journal-title":"IEEE\/CVF International Conference on Computer Vision (ICCV)"},{"key":"2780_CR16","doi-asserted-by":"crossref","unstructured":"Botashev, K., Pyatov, V., Ferrer, G., & Lefkimmiatis, S. (2024). Gsloc: Visual localization with 3d gaussian splatting. In: 2024 IEEE\/RSJ International Conference on Intelligent Robots and Systems (IROS), pp. 5664\u20135671. IEEE.","DOI":"10.1109\/IROS58592.2024.10801919"},{"key":"2780_CR17","doi-asserted-by":"crossref","unstructured":"Brachmann, E., & Rother, C. (2018). Learning Less is More - 6D Camera Localization via 3D Surface Regression. In: CVPR.","DOI":"10.1109\/CVPR.2018.00489"},{"key":"2780_CR18","doi-asserted-by":"crossref","unstructured":"Brachmann, E., Krull, A., Nowozin, S., Shotton, J., Michel, F., Gumhold, S., & Rother, C. (2017). DSAC - Differentiable RANSAC for Camera Localization. In: CVPR.","DOI":"10.1109\/CVPR.2017.267"},{"key":"2780_CR19","first-page":"5044","volume":"2023","author":"E Brachmann","year":"2023","unstructured":"Brachmann, E., Cavallari, T., & Prisacariu, V. A. (2023). Accelerated Coordinate Encoding: Learning to Relocalize in Minutes Using RGB and Poses. IEEE\/CVF Conference on Computer Vision and Pattern Recognition (CVPR), 2023, 5044\u20135053.","journal-title":"IEEE\/CVF Conference on Computer Vision and Pattern Recognition (CVPR)"},{"key":"2780_CR20","first-page":"6198","volume":"2021","author":"E Brachmann","year":"2021","unstructured":"Brachmann, E., Humenberger, M., Rother, C., & Sattler, T. (2021). On the Limits of Pseudo Ground Truth in Visual Camera Re-localisation. IEEE\/CVF International Conference on Computer Vision (ICCV), 2021, 6198\u20136208.","journal-title":"IEEE\/CVF International Conference on Computer Vision (ICCV)"},{"key":"2780_CR21","first-page":"5847","volume":"44","author":"E Brachmann","year":"2020","unstructured":"Brachmann, E., & Rother, C. (2020). Visual Camera Re-Localization From RGB and RGB-D Images Using DSAC. IEEE Transactions on Pattern Analysis and Machine Intelligence, 44, 5847\u20135865.","journal-title":"IEEE Transactions on Pattern Analysis and Machine Intelligence"},{"key":"2780_CR22","unstructured":"Bradski, G. (2000). The OpenCV Library. Dr. Dobb\u2019s Journal of Software Tools."},{"key":"2780_CR23","doi-asserted-by":"crossref","unstructured":"Brahmbhatt, S., Gu, J., Kim, K., Hays, J., & Kautz, J. (2018). Geometry-aware learning of maps for camera localization. In: CVPR.","DOI":"10.1109\/CVPR.2018.00277"},{"key":"2780_CR24","unstructured":"Budvytis, I., Teichmann, M., Vojir, T., & Cipolla, R. (2019). Large Scale Joint Semantic Re-Localisation and Scene Understanding via Globally Unique Instance Coordinate Regression. In: BMVC."},{"key":"2780_CR25","doi-asserted-by":"crossref","unstructured":"Cao, S., & Snavely, N. (2013). Graph-Based Discriminative Learning for Location Recognition. In: CVPR.","DOI":"10.1109\/CVPR.2013.96"},{"key":"2780_CR26","doi-asserted-by":"crossref","unstructured":"Cavallari, T., Bertinetto, L., Mukhoti, J., Torr, P., & Golodetz, S. (2019). Let\u2019s take this online: Adapting scene coordinate regression network predictions for online RGB-D camera relocalisation. In: 3DV.","DOI":"10.1109\/3DV.2019.00068"},{"key":"2780_CR27","doi-asserted-by":"crossref","unstructured":"Cavallari, T., Golodetz, S., Lord, N. A., Valentin, J., Prisacariu, V. A., Di Stefano, L., & Torr, P. H. S. (2019). Real-time RGB-D camera pose estimation in novel scenes using a relocalisation cascade. TPAMI.","DOI":"10.1109\/TPAMI.2019.2915068"},{"key":"2780_CR28","doi-asserted-by":"crossref","unstructured":"Chen, D.M., Baatz, G., K\u00f6ser, K., Tsai, S.S., Vedantham, R., Pylv\u00e4n\u00e4inen, T., Roimela, K., Chen, X., Bach, J., Pollefeys, M., Girod, B., & Grzeszczuk, R. (2011). City-Scale Landmark Identification on Mobile Devices. In: CVPR.","DOI":"10.1109\/CVPR.2011.5995610"},{"key":"2780_CR29","doi-asserted-by":"crossref","unstructured":"Chen, S., Bhalgat, Y., Li, X., Bian, J., Li, K., Wang, Z., & Prisacariu, V.A. (2023). Refinement for absolute pose regression with neural feature synthesis. arXiv preprint arXiv:2303.10087.","DOI":"10.1109\/CVPR52733.2024.01983"},{"key":"2780_CR30","doi-asserted-by":"crossref","unstructured":"Chen, Z., Jacobson, A., S\u00fcnderhauf, N., Upcroft, B., Liu, L., Shen, C., Reid, I.D., & Milford, M. (2017). Deep Learning Features at Scale for Visual Place Recognition. ICRA.","DOI":"10.1109\/ICRA.2017.7989366"},{"key":"2780_CR31","doi-asserted-by":"crossref","unstructured":"Chen, S., Li, X., Wang, Z., & Prisacariu, V.A. (2022). Dfnet: Enhance absolute pose regression with direct feature matching. In: European Conference on Computer Vision, pp. 1\u201317. Springer Nature Switzerland Cham.","DOI":"10.1007\/978-3-031-20080-9_1"},{"key":"2780_CR32","doi-asserted-by":"crossref","unstructured":"Choudhary, S., & Narayanan, P.J. (2012). Visibility probability structure from sfm datasets and applications. In: ECCV.","DOI":"10.1007\/978-3-642-33715-4_10"},{"issue":"8","key":"2780_CR33","doi-asserted-by":"publisher","first-page":"1472","DOI":"10.1109\/TPAMI.2007.70787","volume":"30","author":"O Chum","year":"2008","unstructured":"Chum, O., & Matas, J. (2008). Optimal randomized ransac. IEEE Transactions on Pattern Analysis and Machine Intelligence, 30(8), 1472\u20131482.","journal-title":"IEEE Transactions on Pattern Analysis and Machine Intelligence"},{"key":"2780_CR34","doi-asserted-by":"crossref","unstructured":"Cui, Z., & Tan, P. (2015). Global structure-from-motion by similarity averaging. In: Proceedings of the IEEE International Conference on Computer Vision, pp. 864\u2013872.","DOI":"10.1109\/ICCV.2015.105"},{"key":"2780_CR35","first-page":"337","volume":"2018","author":"D DeTone","year":"2017","unstructured":"DeTone, D., Malisiewicz, T., & Rabinovich, A. (2017). SuperPoint: Self-Supervised Interest Point Detection and Description. IEEE\/CVF Conference on Computer Vision and Pattern Recognition Workshops (CVPRW), 2018, 337\u201333712.","journal-title":"IEEE\/CVF Conference on Computer Vision and Pattern Recognition Workshops (CVPRW)"},{"key":"2780_CR36","doi-asserted-by":"crossref","unstructured":"Ding, Y., Kocur, V., V\u00e1vra, V., Haladov\u00e1, Z.B., Yang, J., Sattler, T., & Kukelova, Z. (2025). RePoseD: Efficient Relative Pose Estimation With Known Depth Information. arXiv preprint arXiv:2501.07742.","DOI":"10.1109\/ICCV51701.2025.01380"},{"key":"2780_CR37","doi-asserted-by":"crossref","unstructured":"Ding, M., Wang, Z., Sun, J., Shi, J., & Luo, P. (2019). CamNet: Coarse-to-fine retrieval for camera re-localization. In: ICCV.","DOI":"10.1109\/ICCV.2019.00296"},{"key":"2780_CR38","doi-asserted-by":"crossref","unstructured":"Ding, Y., Yang, J., Larsson, V., Olsson, C., & \u00c5str\u00f6m, K. (2023). Revisiting the P3P problem. In: Proceedings of the IEEE\/CVF Conference on Computer Vision and Pattern Recognition, pp. 4872\u20134880.","DOI":"10.1109\/CVPR52729.2023.00472"},{"key":"2780_CR39","unstructured":"Dong, S., Liu, S., Guo, H., Chen, B., & Pollefeys, M. (2023). Lazy Visual Localization via Motion Averaging. arXiv:2307.09981."},{"key":"2780_CR40","doi-asserted-by":"crossref","unstructured":"Dong, S., Wang, S., Zhuang, Y., Kannala, J., Pollefeys, M., & Chen, B. (2022). Visual localization via few-shot scene region classification. In: 2022 International Conference on 3D Vision (3DV), pp. 393\u2013402. IEEE.","DOI":"10.1109\/3DV57658.2022.00051"},{"key":"2780_CR41","first-page":"16739","volume":"2025","author":"S Dong","year":"2024","unstructured":"Dong, S., Wang, S., Liu, S., Cai, L., Fan, Q., Kannala, J., & Yang, Y. (2024). Reloc3r: Large-scale training of relative camera pose regression for generalizable, fast, and accurate visual localization. IEEE\/CVF Conference on Computer Vision and Pattern Recognition (CVPR), 2025, 16739\u201316752.","journal-title":"IEEE\/CVF Conference on Computer Vision and Pattern Recognition (CVPR)"},{"key":"2780_CR42","unstructured":"Douze, M., Guzhva, A., Deng, C., Johnson, J., Szilvasy, G., Mazar\u00e9, P.-E., Lomeli, M., Hosseini, L., & J\u00e9gou, H. (2024). The Faiss library arXiv:2401.08281 [cs.LG]."},{"key":"2780_CR43","doi-asserted-by":"crossref","unstructured":"Edstedt, J., Sun, Q., B\u00f6kman, G., Wadenb\u00e4ck, M., & Felsberg, M. (2024). RoMa: Robust Dense Feature Matching. IEEE Conference on Computer Vision and Pattern Recognition.","DOI":"10.1109\/CVPR52733.2024.01871"},{"key":"2780_CR44","doi-asserted-by":"publisher","first-page":"381","DOI":"10.1145\/358669.358692","volume":"24","author":"MA Fischler","year":"1981","unstructured":"Fischler, M. A., & Bolles, R. C. (1981). Random sample consensus: a paradigm for model fitting with applications to image analysis and automated cartography. Commun. ACM, 24, 381\u2013395.","journal-title":"Commun. ACM"},{"key":"2780_CR45","doi-asserted-by":"crossref","unstructured":"Gordo, A., Almazan, J., Revaud, J., & Larlus, D. (2017). End-to-end Learning of Deep Visual Representations for Image Retrieval. IJCV.","DOI":"10.1007\/s11263-017-1016-8"},{"key":"2780_CR46","unstructured":"Grunert, J. A. (1841). Das pothenotische problem in erweiterter gestalt nebst bber seine anwendungen in der geodasie. Grunerts Archiv fur Mathematik und Physik, 238\u2013248."},{"key":"2780_CR47","doi-asserted-by":"crossref","unstructured":"Guzman-Rivera, A., Kohli, P., Glocker, B., Shotton, J., Sharp, T., Fitzgibbon, A., & Izadi, S. (2014). Multi-output learning for camera relocalization. In: CVPR.","DOI":"10.1109\/CVPR.2014.146"},{"key":"2780_CR48","doi-asserted-by":"publisher","first-page":"331","DOI":"10.1007\/BF02028352","volume":"13","author":"BM Haralick","year":"1994","unstructured":"Haralick, B. M., Lee, C.-N., Ottenberg, K., & N\u00f6lle, M. (1994). Review and analysis of solutions of the three point perspective pose estimation problem. International journal of computer vision (IJCV), 13, 331\u2013356.","journal-title":"International journal of computer vision (IJCV)"},{"key":"2780_CR49","unstructured":"Hartley, R., & Zisserman, A. (2001). Multiple View Geometry in Computer Vision, 2nd edn. Cambridge University Press, ???."},{"key":"2780_CR50","doi-asserted-by":"crossref","unstructured":"Hausler, S., Garg, S., Xu, M., Milford, M., & Fischer, T. (2021). Patch-netvlad: Multi-scale fusion of locally-global descriptors for place recognition. In: Proceedings of the IEEE\/CVF Conference on Computer Vision and Pattern Recognition (CVPR), pp. 14141\u201314152.","DOI":"10.1109\/CVPR46437.2021.01392"},{"key":"2780_CR51","first-page":"4695","volume":"2019","author":"L Heng","year":"2018","unstructured":"Heng, L., Choi, B., Cui, Z., Geppert, M., Hu, S., Kuan, B., Liu, P., Nguyen, R. H. M., Yeo, Y. C., Geiger, A., Lee, G. H., Pollefeys, M., & Sattler, T. (2018). Project AutoVision: Localization and 3D Scene Perception for an Autonomous Vehicle with a Multi-Camera System. International Conference on Robotics and Automation (ICRA), 2019, 4695\u20134702.","journal-title":"International Conference on Robotics and Automation (ICRA)"},{"key":"2780_CR52","doi-asserted-by":"crossref","unstructured":"Hu, M., Yin, W., Zhang, C.X., Cai, Z., Long, X., Chen, H., Wang, K., Yu, G., Shen, C., & Shen, S. (2024). Metric3d v2: A versatile monocular geometric foundation model for zero-shot metric depth and surface normal estimation. IEEE transactions on pattern analysis and machine intelligence PP.","DOI":"10.1109\/TPAMI.2024.3444912"},{"key":"2780_CR53","unstructured":"Humenberger, M., Cabon, Y., Guerin, N., Morat, J., Revaud, J., Rerole, P., Pion, N., Souza, C., Leroy, V., & Csurka, G. (2022). Robust Image Retrieval-based Visual Localization using Kapture. arXiv:2007.13867."},{"issue":"7","key":"2780_CR54","doi-asserted-by":"publisher","first-page":"1811","DOI":"10.1007\/s11263-022-01615-7","volume":"130","author":"M Humenberger","year":"2022","unstructured":"Humenberger, M., Cabon, Y., Pion, N., Weinzaepfel, P., Lee, D., Gu\u00e9rin, N., Sattler, T., & Csurka, G. (2022). Investigating the Role of Image Retrieval for Visual Localization: An Exhaustive Benchmark. IJCV, 130(7), 1811\u20131836.","journal-title":"IJCV"},{"key":"2780_CR55","doi-asserted-by":"crossref","unstructured":"Irschara, A., Zach, C., Frahm, J.-M., & Bischof, H. (2009). From Structure-from-Motion Point Clouds to Fast Location Recognition. In: CVPR.","DOI":"10.1109\/CVPR.2009.5206587"},{"key":"2780_CR56","first-page":"3304","volume":"2010","author":"H J\u00e9gou","year":"2010","unstructured":"J\u00e9gou, H., Douze, M., Schmid, C., & P\u00e9rez, P. (2010). Aggregating local descriptors into a compact image representation. IEEE Computer Society Conference on Computer Vision and Pattern Recognition, 2010, 3304\u20133311.","journal-title":"IEEE Computer Society Conference on Computer Vision and Pattern Recognition"},{"key":"2780_CR57","doi-asserted-by":"publisher","DOI":"10.1107\/S0567739476001873","volume-title":"A solution for the best rotation to relate two sets of vectors","author":"W Kabsch","year":"1976","unstructured":"Kabsch, W. (1976). A solution for the best rotation to relate two sets of vectors. Acta Crystallographica Section A: Crystal Physics, Diffraction, Theoretical and General Crystallography."},{"key":"2780_CR58","doi-asserted-by":"crossref","unstructured":"Kazhdan, M., & Hoppe, H. (2013). Screened Poisson Surface Reconstruction. ACM Trans. Graph. 32(3).","DOI":"10.1145\/2487228.2487237"},{"key":"2780_CR59","doi-asserted-by":"crossref","unstructured":"Kendall, A., & Cipolla, R. (2017). Geometric loss functions for camera pose regression with deep learning. In: CVPR.","DOI":"10.1109\/CVPR.2017.694"},{"key":"2780_CR60","doi-asserted-by":"crossref","unstructured":"Kendall, A., Grimes, M., & Cipolla, R. (2015). PoseNet: A Convolutional Network for Real-Time 6-DOF Camera Relocalization. In: ICCV.","DOI":"10.1109\/ICCV.2015.336"},{"issue":"4","key":"2780_CR61","doi-asserted-by":"publisher","first-page":"139","DOI":"10.1145\/3592433","volume":"42","author":"B Kerbl","year":"2023","unstructured":"Kerbl, B., Kopanas, G., Leimk\u00fchler, T., & Drettakis, G. (2023). 3D Gaussian Splatting for Real-Time Radiance Field Rendering. ACM Trans. Graph., 42(4), 139\u20131.","journal-title":"ACM Trans. Graph."},{"key":"2780_CR62","unstructured":"Larsson, V. (2020). contributors: PoseLib - Minimal Solvers for Camera Pose Estimation. https:\/\/github.com\/vlarsson\/PoseLib."},{"key":"2780_CR63","doi-asserted-by":"crossref","unstructured":"Laskar, Z., Melekhov, I., Kalia, S., & Kannala, J. (2017). Camera Relocalization by Computing Pairwise Relative Poses Using Convolutional Neural Network. In: ICCV Workshops.","DOI":"10.1109\/ICCVW.2017.113"},{"key":"2780_CR64","doi-asserted-by":"crossref","unstructured":"Lebeda, K., Matas, J.E.S., & Chum, O. (2012). Fixing the Locally Optimized RANSAC. In: BMVC.","DOI":"10.5244\/C.26.95"},{"key":"2780_CR65","first-page":"564","volume":"2013","author":"GH Lee","year":"2013","unstructured":"Lee, G. H., Fraundorfer, F., & Pollefeys, M. (2013). Structureless Pose-Graph Loop-Closure with a Multi-Camera System on a Self-Driving Car. IEEE\/RSJ International Conference on Intelligent Robots and Systems, 2013, 564\u2013571.","journal-title":"IEEE\/RSJ International Conference on Intelligent Robots and Systems"},{"key":"2780_CR66","first-page":"3226","volume":"2021","author":"D Lee","year":"2021","unstructured":"Lee, D., Ryu, S., Yeon, S., Lee, Y., Kim, D.-W., Han, C., Cabon, Y., Weinzaepfel, P., Gu\u2019erin, N., Csurka, G., & Humenberger, M. (2021). Large-scale Localization Datasets in Crowded Indoor Spaces. IEEE\/CVF Conference on Computer Vision and Pattern Recognition (CVPR), 2021, 3226\u20133235.","journal-title":"IEEE\/CVF Conference on Computer Vision and Pattern Recognition (CVPR)"},{"key":"2780_CR67","doi-asserted-by":"crossref","unstructured":"Leroy, V., Cabon, Y., & Revaud, J. (2024). Grounding Image Matching in 3D with MASt3R.","DOI":"10.1007\/978-3-031-73220-1_5"},{"key":"2780_CR68","doi-asserted-by":"crossref","unstructured":"Li, Y., Snavely, N., & Huttenlocher, D.P. (2010). Location Recognition using Prioritized Feature Matching. In: ECCV.","DOI":"10.1007\/978-3-642-15552-9_57"},{"key":"2780_CR69","doi-asserted-by":"crossref","unstructured":"Li, Y., Snavely, N., Huttenlocher, D. P., & Fua, P. V. (2012). Worldwide Pose Estimation Using 3D Point Clouds. In: European Conference on Computer Vision.","DOI":"10.1007\/978-3-642-33718-5_2"},{"key":"2780_CR70","doi-asserted-by":"crossref","unstructured":"Li, X., Wang, S., Zhao, Y., Verbeek, J., & Kannala, J. (2020). Hierarchical scene coordinate classification and regression for visual localization. In: CVPR.","DOI":"10.1109\/CVPR42600.2020.01200"},{"key":"2780_CR71","first-page":"1043","volume":"2012","author":"H Lim","year":"2012","unstructured":"Lim, H., Sinha, S. N., Cohen, M. F., & Uyttendaele, M. (2012). Real-time image-based 6-DOF localization in large-scale environments. IEEE Conference on Computer Vision and Pattern Recognition, 2012, 1043\u20131050.","journal-title":"IEEE Conference on Computer Vision and Pattern Recognition"},{"key":"2780_CR72","doi-asserted-by":"crossref","unstructured":"Lin, Y.-C., Florence, P.R., Barron, J.T., Rodriguez, A., Isola, P., & Lin, T.-Y. (2020). iNeRF: Inverting Neural Radiance Fields for Pose Estimation. 2021 IEEE\/RSJ International Conference on Intelligent Robots and Systems (IROS), 1323\u20131330.","DOI":"10.1109\/IROS51168.2021.9636708"},{"key":"2780_CR73","doi-asserted-by":"crossref","unstructured":"Lindenberger, P., Sarlin, P.-E., & Pollefeys, M. (2023). LightGlue: Local Feature Matching at Light Speed. In: ICCV.","DOI":"10.1109\/ICCV51070.2023.01616"},{"key":"2780_CR74","unstructured":"Liu, C., Chen, S., Bhalgat, Y.S., Hu, S., Cheng, M., Wang, Z., Prisacariu, V.A., & Braud, T. (2024). Gs-cpr: Efficient camera pose refinement via 3d gaussian splatting. In: The Thirteenth International Conference on Learning Representations."},{"key":"2780_CR75","doi-asserted-by":"crossref","unstructured":"Liu, J., Nie, Q., Liu, Y., & Wang, C. (2023). Nerf-loc: Visual localization with conditional neural radiance field. 2023 IEEE International Conference on Robotics and Automation (ICRA).","DOI":"10.1109\/ICRA48891.2023.10161420"},{"issue":"1","key":"2780_CR76","doi-asserted-by":"publisher","first-page":"1","DOI":"10.1109\/TRO.2015.2496823","volume":"32","author":"S Lowry","year":"2016","unstructured":"Lowry, S., S\u00fcnderhauf, N., Newman, P., Leonard, J. J., Cox, D., Corke, P., & Milford, M. J. (2016). Visual place recognition: A survey. IEEE Transactions on Robotics, 32(1), 1\u201319.","journal-title":"IEEE Transactions on Robotics"},{"key":"2780_CR77","doi-asserted-by":"crossref","unstructured":"Lynen, S., Sattler, T., Bosse, M., Hesch, J.A., Pollefeys, M., & Siegwart, R. Y. (2015). Get Out of My Lab: Large-scale, Real-Time Visual-Inertial Localization. In: Robotics: Science and Systems.","DOI":"10.15607\/RSS.2015.XI.037"},{"key":"2780_CR78","doi-asserted-by":"crossref","unstructured":"Massiceti, D., Krull, A., Brachmann, E., Rother, C., & Torr, P.H.S. (2017). Random Forests versus Neural Networks - What\u2019s Best for Camera Relocalization? In: ICRA.","DOI":"10.1109\/ICRA.2017.7989598"},{"key":"2780_CR79","doi-asserted-by":"publisher","first-page":"420","DOI":"10.1007\/978-3-031-72943-0_24","volume-title":"Computer Vision - ECCV 2024","author":"B Matteo","year":"2025","unstructured":"Matteo, B., Tsesmelis, T., James, S., Poiesi, F., & Del Bue, A. (2025). 6dgs: 6d pose estimation from a single image and a 3d gaussian splatting model. In A. Leonardis, E. Ricci, S. Roth, O. Russakovsky, T. Sattler, & G. Varol (Eds.), Computer Vision - ECCV 2024 (pp. 420\u2013436). Cham: Springer."},{"key":"2780_CR80","doi-asserted-by":"crossref","unstructured":"Middelberg, S., Sattler, T., Untzelmann, O., & Kobbelt, L. P. (2014). Scalable 6-DOF Localization on Mobile Devices. In: European Conference on Computer Vision.","DOI":"10.1007\/978-3-319-10605-2_18"},{"key":"2780_CR81","doi-asserted-by":"publisher","first-page":"99","DOI":"10.1145\/3503250","volume":"65","author":"B Mildenhall","year":"2020","unstructured":"Mildenhall, B., Srinivasan, P. P., Tancik, M., Barron, J. T., Ramamoorthi, R., & Ng, R. (2020). NeRF: Representing Scenes as Neural Radiance Fields for View Synthesis. Commun. ACM, 65, 99\u2013106.","journal-title":"Commun. ACM"},{"key":"2780_CR82","unstructured":"Moreau, A., Piasco, N., Tsishkou, D., Stanciulescu, B., & La Fortelle, A. (2021). LENS: Localization enhanced by neRF synthesis. In: CoRL."},{"key":"2780_CR83","doi-asserted-by":"crossref","unstructured":"Ng, T., Lopez-Rodriguez, A., Balntas, V., & Mikolajczyk, K. (2022). OoD-Pose: Camera Pose Regression From Out-of-Distribution Synthetic Views. In: 2022 International Conference on 3D Vision (3DV).","DOI":"10.1109\/3DV57658.2022.00082"},{"issue":"6","key":"2780_CR84","doi-asserted-by":"publisher","first-page":"756","DOI":"10.1109\/TPAMI.2004.17","volume":"26","author":"D Nist\u00e9r","year":"2004","unstructured":"Nist\u00e9r, D. (2004). An Efficient Solution to the Five-Point Relative Pose Problem. IEEE Transactions on Pattern Analysis and Machine Intelligence, 26(6), 756\u2013770.","journal-title":"IEEE Transactions on Pattern Analysis and Machine Intelligence"},{"key":"2780_CR85","doi-asserted-by":"publisher","first-page":"756","DOI":"10.1109\/TPAMI.2004.17","volume":"26","author":"D Nist\u00e9r","year":"2004","unstructured":"Nist\u00e9r, D. (2004). An efficient solution to the five-point relative pose problem. IEEE Transactions on Pattern Analysis and Machine Intelligence, 26, 756\u2013770.","journal-title":"IEEE Transactions on Pattern Analysis and Machine Intelligence"},{"key":"2780_CR86","doi-asserted-by":"crossref","unstructured":"Pan, L., Bar\u00e1th, D., Pollefeys, M., & Sch\u00f6nberger, J.L. (2024). Global structure-from-motion revisited. In: European Conference on Computer Vision, pp. 58\u201377. Springer.","DOI":"10.1007\/978-3-031-73661-2_4"},{"key":"2780_CR87","doi-asserted-by":"crossref","unstructured":"Panek, V., Kukelova, Z., & Sattler, T. (2022). Meshloc: Mesh-based visual localization. In: ECCV.","DOI":"10.1007\/978-3-031-20047-2_34"},{"key":"2780_CR88","doi-asserted-by":"crossref","unstructured":"Panek, V., Kukelova, Z., & Sattler, T. (2023). Visual Localization Using Imperfect 3D Models From the Internet. In: Proceedings of the IEEE\/CVF Conference on Computer Vision and Pattern Recognition (CVPR), pp. 13175\u201313186.","DOI":"10.1109\/CVPR52729.2023.01266"},{"key":"2780_CR89","doi-asserted-by":"crossref","unstructured":"Panek, V., Sattler, T., & Kukelova, Z. (2025). Combining Absolute and Semi-Generalized Relative Poses for Visual Localization. In: DAGM German Conference on Pattern Recognition. Springer.","DOI":"10.1007\/978-3-032-12840-9_33"},{"key":"2780_CR90","first-page":"13175","volume":"2023","author":"V Panek","year":"2023","unstructured":"Panek, V., Kukelova, Z., & Sattler, T. (2023). Visual Localization using Imperfect 3D Models from the Internet. IEEE\/CVF Conference on Computer Vision and Pattern Recognition (CVPR), 2023, 13175\u201313186.","journal-title":"IEEE\/CVF Conference on Computer Vision and Pattern Recognition (CVPR)"},{"key":"2780_CR91","doi-asserted-by":"crossref","unstructured":"Persson, M., & Nordberg, K. (2018). Lambda twist: An accurate fast robust perspective three point (p3p) solver. In: European Conference on Computer Vision (ECCV).","DOI":"10.1007\/978-3-030-01225-0_20"},{"key":"2780_CR92","doi-asserted-by":"crossref","unstructured":"Philbin, J., Chum, O., Isard, M., Sivic, J., & Zisserman, A. (2007). Object Retrieval with Large Vocabularies and Fast Spatial Matching. In: CVPR.","DOI":"10.1109\/CVPR.2007.383172"},{"key":"2780_CR93","doi-asserted-by":"crossref","unstructured":"Philbin, J., Isard, M., Sivic, J., & Zisserman, A. (2010). Descriptor learning for efficient retrieval. In: ECCV.","DOI":"10.1007\/978-3-642-15558-1_49"},{"key":"2780_CR94","doi-asserted-by":"crossref","unstructured":"Pietrantoni, M., Csurka, G., Humenberger, M., & Sattler, T. (2024). Self-Supervised Learning of Neural Implicit Feature Fields for Camera Pose Refinement. In: 2024 International Conference on 3D Vision (3DV), pp. 484\u2013494. IEEE.","DOI":"10.1109\/3DV62453.2024.00139"},{"key":"2780_CR95","doi-asserted-by":"crossref","unstructured":"Pietrantoni, M., Humenberger, M., Sattler, T., & Csurka, G. (2023). Segloc: Learning segmentation-based representations for privacy-preserving visual localization. In: Proceedings of the IEEE\/CVF Conference on Computer Vision and Pattern Recognition (CVPR), pp. 15380\u201315391.","DOI":"10.1109\/CVPR52729.2023.01476"},{"key":"2780_CR96","doi-asserted-by":"crossref","unstructured":"Pion, N., Humenberger, M., Csurka, G., Cabon, Y., & Sattler, T. (2020). Benchmarking Image Retrieval for Visual Localization. In: 3DV.","DOI":"10.1109\/3DV50981.2020.00058"},{"key":"2780_CR97","doi-asserted-by":"crossref","unstructured":"Radenovic, F., Tolias, G., & Chum, O. (2019) Fine-Tuning CNN Image Retrieval with No Human Annotation. TPAMI.","DOI":"10.1109\/TPAMI.2018.2846566"},{"key":"2780_CR98","doi-asserted-by":"crossref","unstructured":"Revaud, J., Almazan, J., Rezende, R.S., & Souza, C.R. (2019). Learning with Average Precision: Training Image Retrieval with a Listwise Loss. In: ICCV.","DOI":"10.1109\/ICCV.2019.00521"},{"key":"2780_CR99","unstructured":"Revaud, J., Weinzaepfel, P., Souza, C.R., & Humenberger, M. (2019). R2D2: repeatable and reliable detector and descriptor. In: NeurIPS."},{"key":"2780_CR100","doi-asserted-by":"crossref","unstructured":"Sarlin, P.-E., Cadena, C., Siegwart, R., & Dymczyk, M. (2019). From Coarse to Fine: Robust Hierarchical Localization at Large Scale. In: CVPR.","DOI":"10.1109\/CVPR.2019.01300"},{"key":"2780_CR101","unstructured":"Sarlin, P.-E., Debraine, F., Dymczyk, M., Siegwart, R., & Cadena, C. (2018). Leveraging Deep Visual Descriptors for Hierarchical Efficient Localization. In: Conference on Robot Learning (CoRL)."},{"key":"2780_CR102","doi-asserted-by":"crossref","unstructured":"Sarlin, P.-E., DeTone, D., Malisiewicz, T., & Rabinovich, A. (2020). SuperGlue: Learning Feature Matching with Graph Neural Networks. In: CVPR.","DOI":"10.1109\/CVPR42600.2020.00499"},{"key":"2780_CR103","doi-asserted-by":"crossref","unstructured":"Sarlin, P.-E., Unagar, A., Larsson, M., Germain, H., Toft, C., Larsson, V., Pollefeys, M., Lepetit, V., Hammarstrand, L., Kahl, F., & Sattler, T. (2021). Back to the Feature: Learning Robust Camera Localization from Pixels to Pose. In: CVPR.","DOI":"10.1109\/CVPR46437.2021.00326"},{"key":"2780_CR104","doi-asserted-by":"crossref","unstructured":"Sattler, T., Havlena, M., Radenovic, F., Schindler, K., & Pollefeys, M. (2015). Hyperpoints and fine vocabularies for large-scale location recognition. In: ICCV.","DOI":"10.1109\/ICCV.2015.243"},{"key":"2780_CR105","doi-asserted-by":"crossref","unstructured":"Sattler, T., Leibe, B., & Kobbelt, L. (2011). Fast Image-Based Localization using Direct 2D-to-3D Matching. In: ICCV.","DOI":"10.1109\/ICCV.2011.6126302"},{"key":"2780_CR106","doi-asserted-by":"crossref","unstructured":"Sattler, T., Leibe, B., & Kobbelt, L. (2012). Improving Image-Based Localization by Active Correspondence Search. In: ECCV.","DOI":"10.1007\/978-3-642-33718-5_54"},{"key":"2780_CR107","doi-asserted-by":"crossref","unstructured":"Sattler, T., Maddern, W., Toft, C., Torii, A., Hammarstrand, L., Stenborg, E., Safari, D., Okutomi, M., Pollefeys, M., Sivic, J., Kahl, F., & Pajdla, T. (2018). Benchmarking 6DOF Urban Visual Localization in Changing Conditions. In: CVPR.","DOI":"10.1109\/CVPR.2018.00897"},{"key":"2780_CR108","doi-asserted-by":"crossref","unstructured":"Sattler, T., Weyand, T., Leibe, B., & Kobbelt, L. (2012). Image Retrieval for Image-Based Localization Revisited. In: BMVC.","DOI":"10.5244\/C.26.76"},{"key":"2780_CR109","doi-asserted-by":"crossref","unstructured":"Sattler, T., Zhou, Q., Pollefeys, M., & Leal-Taix\u00e9, L. (2019). Understanding the limitations of cnn-based absolute camera pose regression. In: CVPR.","DOI":"10.1109\/CVPR.2019.00342"},{"key":"2780_CR110","doi-asserted-by":"publisher","first-page":"1744","DOI":"10.1109\/TPAMI.2016.2611662","volume":"39","author":"T Sattler","year":"2017","unstructured":"Sattler, T., Leibe, B., & Kobbelt, L. P. (2017). Efficient & Effective Prioritized Matching for Large-Scale Image-Based Localization. IEEE Transactions on Pattern Analysis and Machine Intelligence, 39, 1744\u20131756.","journal-title":"IEEE Transactions on Pattern Analysis and Machine Intelligence"},{"key":"2780_CR111","unstructured":"Se, S., Lowe, D., & Little, J. (2002). Global localization using distinctive visual features. In: IEEE\/RSJ International Conference on Intelligent Robots and Systems."},{"key":"2780_CR112","doi-asserted-by":"crossref","unstructured":"Shavit, Y., Ferens, R., & Keller, Y. (2021). Learning Multi-Scene Absolute Pose Regression With Transformers. In: ICCV.","DOI":"10.1109\/ICCV48922.2021.00273"},{"key":"2780_CR113","first-page":"2930","volume":"2013","author":"J Shotton","year":"2013","unstructured":"Shotton, J., Glocker, B., Zach, C., Izadi, S., Criminisi, A., & Fitzgibbon, A. W. (2013). Scene Coordinate Regression Forests for Camera Relocalization in RGB-D Images. IEEE Conference on Computer Vision and Pattern Recognition, 2013, 2930\u20132937.","journal-title":"IEEE Conference on Computer Vision and Pattern Recognition"},{"key":"2780_CR114","doi-asserted-by":"crossref","unstructured":"Singh, G., & Ko\u0161eck\u00e1, J. (2016). Semantically Guided Geo-location and Modeling in Urban Environments. In: Large-Scale Visual Geo-Localization.","DOI":"10.1007\/978-3-319-25781-5_6"},{"key":"2780_CR115","doi-asserted-by":"crossref","unstructured":"Sivic, J., & Zisserman, A. (2003). Video Google: A Text Retrieval Approach to Object Matching in Videos. In: ICCV.","DOI":"10.1109\/ICCV.2003.1238663"},{"key":"2780_CR116","unstructured":"Sun, Y., Wang, X., Zhang, Y., Zhang, J., Jiang, C., Guo, Y., & Wang, F. (2023). icomma: Inverting 3d gaussian splatting for camera pose estimation via comparing and matching. arXiv preprint arXiv:2312.09031."},{"key":"2780_CR117","first-page":"8918","volume":"2021","author":"J Sun","year":"2021","unstructured":"Sun, J., Shen, Z., Wang, Y., Bao, H., & Zhou, X. (2021). LoFTR: Detector-Free Local Feature Matching with Transformers. IEEE\/CVF Conference on Computer Vision and Pattern Recognition (CVPR), 2021, 8918\u20138927.","journal-title":"IEEE\/CVF Conference on Computer Vision and Pattern Recognition (CVPR)"},{"issue":"7","key":"2780_CR118","doi-asserted-by":"publisher","first-page":"1455","DOI":"10.1109\/TPAMI.2016.2598331","volume":"39","author":"L Sv\u00e4rm","year":"2017","unstructured":"Sv\u00e4rm, L., Enqvist, O., Kahl, F., & Oskarsson, M. (2017). City-Scale Localization for Cameras with Known Vertical Direction. PAMI, 39(7), 1455\u20131461.","journal-title":"PAMI"},{"key":"2780_CR119","doi-asserted-by":"crossref","unstructured":"Taira, H., Okutomi, M., Sattler, T., Cimpoi, M., Pollefeys, M., Sivic, J., Pajdla, T., & Torii, A. (2021). InLoc: Indoor Visual Localization with Dense Matching and View Synthesis. TPAMI.","DOI":"10.1109\/TPAMI.2019.2952114"},{"key":"2780_CR120","doi-asserted-by":"publisher","first-page":"1293","DOI":"10.1109\/TPAMI.2019.2952114","volume":"43","author":"H Taira","year":"2019","unstructured":"Taira, H., Okutomi, M., Sattler, T., Cimpoi, M., Pollefeys, M., Sivic, J., Pajdla, T., & Torii, A. (2019). InLoc: Indoor Visual Localization with Dense Matching and View Synthesis. IEEE Transactions on Pattern Analysis and Machine Intelligence, 43, 1293\u20131307.","journal-title":"IEEE Transactions on Pattern Analysis and Machine Intelligence"},{"key":"2780_CR121","doi-asserted-by":"crossref","unstructured":"Tang, S., Tang, C., Huang, R., Zhu, S., & Tan, P. (2021). Learning camera localization via dense scene matching. In: Proceedings of the IEEE\/CVF Conference on Computer Vision and Pattern Recognition, pp. 1831\u20131841.","DOI":"10.1109\/CVPR46437.2021.00187"},{"issue":"2","key":"2780_CR122","doi-asserted-by":"publisher","first-page":"2062","DOI":"10.1109\/LRA.2020.2970616","volume":"5","author":"J Thoma","year":"2020","unstructured":"Thoma, J., Paudel, D. P., Chhatkuli, A., & Gool, L. V. (2020). Geometrically mappable image features. IEEE Robotics and Automation Letters, 5(2), 2062\u20132069.","journal-title":"IEEE Robotics and Automation Letters"},{"issue":"3","key":"2780_CR123","doi-asserted-by":"publisher","first-page":"247","DOI":"10.1007\/s11263-015-0810-4","volume":"116","author":"G Tolias","year":"2016","unstructured":"Tolias, G., Avrithis, Y. S., & J\u00e9gou, H. (2016). Image Search with Selective Match Kernels: Aggregation Across Single and Multiple Images. IJCV, 116(3), 247\u2013261.","journal-title":"IJCV"},{"key":"2780_CR124","doi-asserted-by":"crossref","unstructured":"Torii, A., Arandjelovi\u0107, R., Sivic, J., Okutomi, M., & Pajdla, T. (2015). 24\/7 place recognition by view synthesis. In: CVPR.","DOI":"10.1109\/CVPR.2015.7298790"},{"key":"2780_CR125","doi-asserted-by":"crossref","unstructured":"Torii, A., Taira, H., Sivic, J., Pollefeys, M., Okutomi, M., Pajdla, T., & Sattler, T. (2021). Are Large-Scale 3D Models Really Necessary for Accurate Visual Localization? TPAMI.","DOI":"10.1109\/TPAMI.2019.2941876"},{"key":"2780_CR126","doi-asserted-by":"crossref","unstructured":"Trivigno, G., Masone, C., Caputo, B., & Sattler, T. (2024). The Unreasonable Effectiveness of Pre-Trained Features for Camera Pose Refinement. In: Proceedings of the IEEE\/CVF Conference on Computer Vision and Pattern Recognition (CVPR), pp. 12786\u201312798.","DOI":"10.1109\/CVPR52733.2024.01215"},{"issue":"04","key":"2780_CR127","doi-asserted-by":"publisher","first-page":"376","DOI":"10.1109\/34.88573","volume":"13","author":"S Umeyama","year":"1991","unstructured":"Umeyama, S. (1991). Least-Squares Estimation of Transformation Parameters Between Two Point Patterns. IEEE Transactions on Pattern Analysis & Machine Intelligence, 13(04), 376\u2013380.","journal-title":"IEEE Transactions on Pattern Analysis & Machine Intelligence"},{"key":"2780_CR128","doi-asserted-by":"crossref","unstructured":"Valentin, J., Nie\u00dfner, M., Shotton, J., Fitzgibbon, A., Izadi, S., & Torr, P. (2015). Exploiting Uncertainty in Regression Forests for Accurate Camera Relocalization. In: CVPR.","DOI":"10.1109\/CVPR.2015.7299069"},{"key":"2780_CR129","doi-asserted-by":"crossref","unstructured":"Von Stumberg, L., Wenzel, P., Yang, N., & Cremers, D. (2020). LM-Reloc: Levenberg-Marquardt Based Direct Visual Relocalization. In: 2020 International Conference on 3D Vision (3DV), pp. 968\u2013977. IEEE.","DOI":"10.1109\/3DV50981.2020.00107"},{"issue":"2","key":"2780_CR130","doi-asserted-by":"publisher","first-page":"890","DOI":"10.1109\/LRA.2020.2965031","volume":"5","author":"L Von Stumberg","year":"2020","unstructured":"Von Stumberg, L., Wenzel, P., Khan, Q., & Cremers, D. (2020). GN-Net: The Gauss-Newton Loss for Multi-Weather Relocalization. IEEE Robotics and Automation Letters, 5(2), 890\u2013897.","journal-title":"IEEE Robotics and Automation Letters"},{"key":"2780_CR131","doi-asserted-by":"crossref","unstructured":"Walch, F., Hazirbas, C., Leal-Taix\u00e9, L., Sattler, T., Hilsenbeck, S., & Cremers, D. (2017). Image-Based Localization Using LSTMs for Structured Feature Correlation. In: ICCV.","DOI":"10.1109\/ICCV.2017.75"},{"key":"2780_CR132","doi-asserted-by":"crossref","unstructured":"Wang, J., Chen, M., Karaev, N., Vedaldi, A., Rupprecht, C., & Novotny, D. (2025). VGGT: Visual Geometry Grounded Transformer. In: Proceedings of the IEEE\/CVF Conference on Computer Vision and Pattern Recognition.","DOI":"10.1109\/CVPR52734.2025.00499"},{"key":"2780_CR133","doi-asserted-by":"crossref","unstructured":"Wang, S., Kannala, J., & Barath, D. (2024). Dgc-gnn: leveraging geometry and color cues for visual descriptor-free 2d\u20133d matching. In: Proceedings of the IEEE\/CVF Conference on Computer Vision and Pattern Recognition, pp. 20881\u201320891.","DOI":"10.1109\/CVPR52733.2024.01973"},{"key":"2780_CR134","doi-asserted-by":"crossref","unstructured":"Wang, S., Laskar, Z., Melekhov, I., Li, X., Zhao, Y., Tolias, G., & Kannala, J. (2023). HSCNet++: Hierarchical Scene Coordinate Classification and Regression for Visual Localization with Transformer. arXiv:2305.03595 [cs.CV].","DOI":"10.1007\/s11263-023-01982-9"},{"key":"2780_CR135","doi-asserted-by":"crossref","unstructured":"Wang, S., Leroy, V., Cabon, Y., Chidlovskii, B., & Revaud, J. (2024). DUSt3R: Geometric 3D Vision Made Easy. In: CVPR.","DOI":"10.1109\/CVPR52733.2024.01956"},{"key":"2780_CR136","doi-asserted-by":"crossref","unstructured":"Yew, Z.J., & Lee, G.H. (2021). City-scale scene change detection using point clouds. In: 2021 IEEE International Conference on Robotics and Automation (ICRA), pp. 13362\u201313369. IEEE.","DOI":"10.1109\/ICRA48506.2021.9561855"},{"key":"2780_CR137","doi-asserted-by":"crossref","unstructured":"Yin, W., Zhang, C., Chen, H., Cai, Z., Yu, G., Wang, K., Chen, X., & Shen, C. (2023). Metric3d: Towards zero-shot metric 3d prediction from a single image. In: ICCV.","DOI":"10.1109\/ICCV51070.2023.00830"},{"key":"2780_CR138","doi-asserted-by":"crossref","unstructured":"Zamir, A.R., & Shah, M. (2010). Accurate Image Localization Based on Google Maps Street View. In: ECCV.","DOI":"10.1007\/978-3-642-15561-1_19"},{"issue":"8","key":"2780_CR139","doi-asserted-by":"publisher","first-page":"1546","DOI":"10.1109\/TPAMI.2014.2299799","volume":"36","author":"AR Zamir","year":"2014","unstructured":"Zamir, A. R., & Shah, M. (2014). Image Geo-Localization Based on MultipleNearest Neighbor Feature Matching Using Generalized Graphs. PAMI, 36(8), 1546\u20131558.","journal-title":"PAMI"},{"key":"2780_CR140","doi-asserted-by":"crossref","unstructured":"Zeisl, B., Sattler, T., & Pollefeys, M. (2015). Camera pose voting for large-scale image-based localization. In: ICCV.","DOI":"10.1109\/ICCV.2015.310"},{"key":"2780_CR141","unstructured":"Zeller, A.J., & Wu, H. (2024). Gsplatloc: Ultra-precise camera localization via 3d gaussian splatting. arXiv preprint arXiv:2412.20056."},{"key":"2780_CR142","doi-asserted-by":"crossref","unstructured":"Zhang, W., & Kosecka, J. (2006). Image Based Localization in Urban Environments. Third International Symposium on 3D Data Processing, Visualization, and Transmission (3DPVT\u201906), 33\u201340.","DOI":"10.1109\/3DPVT.2006.80"},{"key":"2780_CR143","unstructured":"Zhang, Z., Sattler, T., & Scaramuzza, D. (2020). Reference Pose Generation for Visual Localization via Learned Features and View Synthesis. arXiv 2005.05179."},{"key":"2780_CR144","doi-asserted-by":"publisher","first-page":"1","DOI":"10.1109\/TIM.2023.3271000","volume":"72","author":"X Zhao","year":"2023","unstructured":"Zhao, X., Wu, X., Chen, W., Chen, P. C. Y., Xu, Q., & Li, Z. (2023). Aliked: A lighter keypoint and descriptor extraction network via deformable transformation. IEEE Transactions on Instrumentation & Measurement, 72, 1\u201316. https:\/\/doi.org\/10.1109\/TIM.2023.3271000","journal-title":"IEEE Transactions on Instrumentation & Measurement"},{"key":"2780_CR145","doi-asserted-by":"publisher","DOI":"10.1109\/TMM.2022.3155927","author":"X Zhao","year":"2022","unstructured":"Zhao, X., Wu, X., Miao, J., Chen, W., Chen, P. C. Y., & Li, Z. (2022). Alike: Accurate and lightweight keypoint detection and descriptor extraction. IEEE Transactions on Multimedia. https:\/\/doi.org\/10.1109\/TMM.2022.3155927","journal-title":"IEEE Transactions on Multimedia"},{"key":"2780_CR146","doi-asserted-by":"crossref","unstructured":"Zheng, E., & Wu, C. (2015). Structure from Motion Using Structure-Less Resection. In: ICCV.","DOI":"10.1109\/ICCV.2015.240"},{"key":"2780_CR147","doi-asserted-by":"crossref","unstructured":"Zhou, Q., Agostinho, S., O\u0161ep, A., & Leal-Taix\u00e9, L. (2022). Is geometry enough for matching in visual localization? In: European Conference on Computer Vision, pp. 407\u2013425. Springer.","DOI":"10.1007\/978-3-031-20080-9_24"},{"key":"2780_CR148","doi-asserted-by":"crossref","unstructured":"Zhou, Q., Maximov, M., Litany, O., & Leal-Taix\u00e9, L. (2024). The nerfect match: Exploring nerf features for visual localization. In: European Conference on Computer Vision, pp. 108\u2013127. Springer.","DOI":"10.1007\/978-3-031-72691-0_7"},{"key":"2780_CR149","doi-asserted-by":"crossref","unstructured":"Zhou, Q., Sattler, T., Pollefeys, M., & Leal-Taix\u00e9, L. (2019). To Learn or Not to Learn: Visual Localization from Essential Matrices. In: ICRA.","DOI":"10.1109\/ICRA40945.2020.9196607"},{"key":"2780_CR150","first-page":"4667","volume":"2021","author":"Q Zhou","year":"2020","unstructured":"Zhou, Q., Sattler, T., & Leal-Taix\u00e9, L. (2020). Patch2Pix: Epipolar-Guided Pixel-Level Correspondences. IEEE\/CVF Conference on Computer Vision and Pattern Recognition (CVPR), 2021, 4667\u20134676.","journal-title":"IEEE\/CVF Conference on Computer Vision and Pattern Recognition (CVPR)"},{"key":"2780_CR151","doi-asserted-by":"crossref","unstructured":"Zhu, S., Zhang, R., Zhou, L., Shen, T., Fang, T., Tan, P., & Quan, L. (2018). Very large-scale global sfm by distributed motion averaging. In: Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition, pp. 4568\u20134577.","DOI":"10.1109\/CVPR.2018.00480"}],"container-title":["International Journal of Computer Vision"],"original-title":[],"language":"en","link":[{"URL":"https:\/\/link.springer.com\/content\/pdf\/10.1007\/s11263-026-02780-9.pdf","content-type":"application\/pdf","content-version":"vor","intended-application":"text-mining"},{"URL":"https:\/\/link.springer.com\/article\/10.1007\/s11263-026-02780-9","content-type":"text\/html","content-version":"vor","intended-application":"text-mining"},{"URL":"https:\/\/link.springer.com\/content\/pdf\/10.1007\/s11263-026-02780-9.pdf","content-type":"application\/pdf","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2026,7,20]],"date-time":"2026-07-20T16:17:45Z","timestamp":1784564265000},"score":1,"resource":{"primary":{"URL":"https:\/\/link.springer.com\/10.1007\/s11263-026-02780-9"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2026,5,9]]},"references-count":151,"journal-issue":{"issue":"6","published-print":{"date-parts":[[2026,6]]}},"alternative-id":["2780"],"URL":"https:\/\/doi.org\/10.1007\/s11263-026-02780-9","relation":{},"ISSN":["0920-5691","1573-1405"],"issn-type":[{"value":"0920-5691","type":"print"},{"value":"1573-1405","type":"electronic"}],"subject":[],"published":{"date-parts":[[2026,5,9]]},"assertion":[{"value":"22 April 2025","order":1,"name":"received","label":"Received","group":{"name":"ArticleHistory","label":"Article History"}},{"value":"5 February 2026","order":2,"name":"accepted","label":"Accepted","group":{"name":"ArticleHistory","label":"Article History"}},{"value":"9 May 2026","order":3,"name":"first_online","label":"First Online","group":{"name":"ArticleHistory","label":"Article History"}}],"article-number":"263"}}