{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2026,7,13]],"date-time":"2026-07-13T20:12:09Z","timestamp":1783973529229,"version":"3.55.0"},"reference-count":38,"publisher":"MDPI AG","issue":"10","license":[{"start":{"date-parts":[[2025,10,13]],"date-time":"2025-10-13T00:00:00Z","timestamp":1760313600000},"content-version":"vor","delay-in-days":0,"URL":"https:\/\/creativecommons.org\/licenses\/by\/4.0\/"}],"funder":[{"DOI":"10.13039\/501100001809","name":"National Natural Science Foundation of China","doi-asserted-by":"crossref","award":["41401436"],"award-info":[{"award-number":["41401436"]}],"id":[{"id":"10.13039\/501100001809","id-type":"DOI","asserted-by":"crossref"}]},{"name":"Innovation and Development Special Project of China Meteorological Administration","award":["CXFZ2025Q009"],"award-info":[{"award-number":["CXFZ2025Q009"]}]},{"DOI":"10.13039\/501100012337","name":"Nanhu Scholars Program for Young Scholars of XYNU","doi-asserted-by":"crossref","id":[{"id":"10.13039\/501100012337","id-type":"DOI","asserted-by":"crossref"}]},{"name":"Key Scientific and Technological Research Project of Henan Province","award":["222102210320"],"award-info":[{"award-number":["222102210320"]}]}],"content-domain":{"domain":[],"crossmark-restriction":false},"short-container-title":["IJGI"],"abstract":"<jats:p>Converting tower-mounted videos from perspective to orthographic view is beneficial for their integration with maps and remote sensing images and can provide a clearer and more real-time data source for earth observation. This paper addresses the issue of low geometric accuracy in orthographic video generation by proposing a method that incorporates 3D GIS view matching. Firstly, a geometric alignment model between video frames and 3D GIS views is established through camera parameter mapping. Then, feature point detection and matching algorithms are employed to associate image coordinates with corresponding 3D spatial coordinates. Finally, an orthographic video map is generated based on the color point cloud. The results show that (1) for tower-based video, a 3D GIS constructed from publicly available DEMs and high-resolution remote sensing imagery can meet the spatialization needs of large-scale tower-mounted video data. (2) The feature point matching algorithm based on deep learning effectively achieves accurate matching between video frames and 3D GIS views. (3) Compared with the traditional method, such as the camera parameters method, the orthographic video map generated by this method has advantages in terms of geometric mapping accuracy and visualization effect. In the mountainous area, the RMSE of the control points is reduced from 137.70 m to 7.72 m. In the flat area, it is reduced from 13.52 m to 8.10 m. The proposed method can provide a near-real-time orthographic video map for smart cities, natural resource monitoring, emergency rescue, and other fields.<\/jats:p>","DOI":"10.3390\/ijgi14100398","type":"journal-article","created":{"date-parts":[[2025,10,14]],"date-time":"2025-10-14T14:34:15Z","timestamp":1760452455000},"page":"398","update-policy":"https:\/\/doi.org\/10.3390\/mdpi_crossmark_policy","source":"Crossref","is-referenced-by-count":1,"title":["Orthographic Video Map Generation Considering 3D GIS View Matching"],"prefix":"10.3390","volume":"14","author":[{"given":"Xingguo","family":"Zhang","sequence":"first","affiliation":[{"name":"School of Geographic Sciences, Xinyang Normal University, Xinyang 464000, China"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Xiangfei","family":"Meng","sequence":"additional","affiliation":[{"name":"School of Geographic Sciences, Xinyang Normal University, Xinyang 464000, China"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Li","family":"Zhang","sequence":"additional","affiliation":[{"name":"School of Physics and Electronic Engineering, Xinyang Normal University, Xinyang 464000, China"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Xianguo","family":"Ling","sequence":"additional","affiliation":[{"name":"School of Geographic Sciences, Xinyang Normal University, Xinyang 464000, China"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Sen","family":"Yang","sequence":"additional","affiliation":[{"name":"School of Geographic Sciences, Xinyang Normal University, Xinyang 464000, China"}],"role":[{"vocabulary":"crossref","role":"author"}]}],"member":"1968","published-online":{"date-parts":[[2025,10,13]]},"reference":[{"key":"ref_1","first-page":"1","article-title":"On geospatial information science in the era of loE","volume":"51","author":"Li","year":"2022","journal-title":"Acta Geod. Cartogr. Sin."},{"key":"ref_2","doi-asserted-by":"crossref","first-page":"1415","DOI":"10.1080\/13658811003792213","article-title":"GIS-augmented video surveillance","volume":"24","year":"2010","journal-title":"Int. J. Geogr. Inf. Sci."},{"key":"ref_3","first-page":"VI-401","article-title":"GPS, GIS and video registration for building reconstruction","volume":"Volume 6","author":"Sourimant","year":"2007","journal-title":"Proceedings of the 2007 IEEE International Conference on Image Processing"},{"key":"ref_4","first-page":"1","article-title":"Research and application on key technologies of natural resources intelligent monitoring with tower-based video","volume":"2023","author":"Chen","year":"2023","journal-title":"Nat. Resour. Informatiz."},{"key":"ref_5","doi-asserted-by":"crossref","unstructured":"Sankaranarayanan, K., and Davis, J.W. (2008, January 1\u20133). A fast linear registration framework for multi-camera GIS coordination. Proceedings of the 2008 IEEE Fifth International Conference on Advanced Video and Signal Based Surveillance, Santa Fe, NM, USA.","DOI":"10.1109\/AVSS.2008.20"},{"key":"ref_6","first-page":"2089","article-title":"Integration of GIS and video surveillance","volume":"30","year":"2016","journal-title":"Int. J. Geogr. Inf. Sci."},{"key":"ref_7","doi-asserted-by":"crossref","first-page":"450","DOI":"10.1111\/tgis.12696","article-title":"Spatiotemporal retrieval of dynamic video object trajectories in geographical scenes","volume":"25","author":"Xie","year":"2021","journal-title":"Trans. GIS"},{"key":"ref_8","doi-asserted-by":"crossref","first-page":"913","DOI":"10.1080\/13658816.2022.2158190","article-title":"Complete trajectory extraction for moving targets in traffic scenes that considers multi-level semantic features","volume":"37","author":"Luo","year":"2023","journal-title":"Int. J. Geogr. Inf. Sci."},{"key":"ref_9","doi-asserted-by":"crossref","unstructured":"Zhang, X., Shi, X., Luo, X., Sun, Y., and Zhou, Y. (2021). Real-time web map construction based on multiple cameras and GIS. ISPRS Int. J. Geo Inf., 10.","DOI":"10.3390\/ijgi10120803"},{"key":"ref_10","doi-asserted-by":"crossref","unstructured":"Shao, Z., Li, C., Li, D., Altan, O., Zhang, L., and Ding, L. (2020). An accurate matching method for projecting vector data into surveillance video to monitor and protect cultivated land. ISPRS Int. J. Geo Inf., 9.","DOI":"10.3390\/ijgi9070448"},{"key":"ref_11","first-page":"1130","article-title":"Mutual Mapping Between Surveillance Video and 2D Geospatial Data","volume":"40","author":"Zhang","year":"2015","journal-title":"Geomat. Inf. Sci. Wuhan Univ."},{"key":"ref_12","doi-asserted-by":"crossref","first-page":"1330","DOI":"10.1109\/34.888718","article-title":"A flexible new technique for camera calibration","volume":"22","author":"Zhang","year":"2002","journal-title":"IEEE Trans. Pattern Anal. Mach. Intell."},{"key":"ref_13","doi-asserted-by":"crossref","first-page":"323","DOI":"10.1109\/JRA.1987.1087109","article-title":"A versatile camera calibration technique for high-accuracy 3D machine vision metrology using off-the-shelf TV cameras and lenses","volume":"3","author":"Tsai","year":"2003","journal-title":"IEEE J. Robot. Autom."},{"key":"ref_14","doi-asserted-by":"crossref","first-page":"5","DOI":"10.1023\/A:1007957826135","article-title":"Self-calibration of stationary cameras","volume":"22","author":"Hartley","year":"1997","journal-title":"Int. J. Comput. Vis."},{"key":"ref_15","first-page":"428","article-title":"An active vision based camera intrinsic parameters self-calibration technique","volume":"21","author":"Yang","year":"1998","journal-title":"Chin. J. Comput."},{"key":"ref_16","first-page":"276","article-title":"Self-Calibration of a Camera with a Non-Linear Model","volume":"25","author":"Hou","year":"2002","journal-title":"Chin. J. Comput."},{"key":"ref_17","doi-asserted-by":"crossref","first-page":"1036","DOI":"10.1109\/34.329005","article-title":"Projective reconstruction and invariants from multiple images","volume":"16","author":"Hartley","year":"1994","journal-title":"IEEE Trans. Pattern Anal. Mach. Intell."},{"key":"ref_18","doi-asserted-by":"crossref","first-page":"349","DOI":"10.1109\/ICPR.1996.546047","article-title":"The modulus constraint: A new constraint self-calibration","volume":"Volume 1","author":"Pollefeys","year":"1996","journal-title":"Proceedings of the 13th International Conference on Pattern Recognition"},{"key":"ref_19","doi-asserted-by":"crossref","first-page":"107","DOI":"10.1023\/A:1012471930694","article-title":"Self-calibration of rotating and zooming cameras","volume":"45","author":"Agapito","year":"2001","journal-title":"Int. J. Comput. Vis."},{"key":"ref_20","unstructured":"Hemayed, E.E. (2003, January 22). A survey of camera self-calibration. Proceedings of the IEEE Conference on Advanced Video and Signal Based Surveillance, Miami, FL, USA."},{"key":"ref_21","unstructured":"Triggs, B. (1997, January 17\u201319). Autocalibration and the absolute quadric. Proceedings of the IEEE Computer Society Conference on Computer Vision and Pattern Recognition, San Juan, PR, USA."},{"key":"ref_22","doi-asserted-by":"crossref","first-page":"350","DOI":"10.1109\/ICCV.1999.791241","article-title":"Flexible calibration: Minimal cases for auto-calibration","volume":"Volume 1","author":"Heyden","year":"1999","journal-title":"Proceedings of the Seventh IEEE International Conference on Computer Vision"},{"key":"ref_23","doi-asserted-by":"crossref","first-page":"510","DOI":"10.1109\/ICCV.1999.791264","article-title":"Camera calibration and the search for infinity","volume":"Volume 1","author":"Hartley","year":"1999","journal-title":"Proceedings of the Seventh IEEE International Conference on Computer Vision"},{"key":"ref_24","first-page":"2243","article-title":"Nonlinear camera model calibrated by neural network and adaptive genetic-annealing algorithm","volume":"27","author":"Ge","year":"2014","journal-title":"J. Intell. Fuzzy Syst."},{"key":"ref_25","doi-asserted-by":"crossref","unstructured":"Woo, D.M., and Park, D.C. (2009, January 3\u20135). An efficient method for camera calibration using multilayer perceptron type neural network. Proceedings of the 2009 International Conference on Future Computer and Communication, Kuala Lumpar, Malaysia.","DOI":"10.1109\/ICFCC.2009.94"},{"key":"ref_26","doi-asserted-by":"crossref","unstructured":"Woo, D.M., and Park, D.C. (2009, January 1\u20133). Implicit camera calibration using MultiLayer perceptron type neural network. Proceedings of the 2009 First Asian Conference on Intelligent Information and Database Systems, Dong hoi, Vietnam.","DOI":"10.1109\/ACIIDS.2009.11"},{"key":"ref_27","doi-asserted-by":"crossref","first-page":"464","DOI":"10.1016\/j.precisioneng.2024.02.019","article-title":"A deep-learning based high-accuracy camera calibration method for large-scale scene","volume":"88","author":"Duan","year":"2024","journal-title":"Precis. Eng."},{"key":"ref_28","doi-asserted-by":"crossref","unstructured":"Li, B., Wang, X., Gao, Q., Song, Z., Zou, C., and Liu, S. (2022). A 3D scene information enhancement method applied in augmented reality. Electronics, 11.","DOI":"10.3390\/electronics11244123"},{"key":"ref_29","first-page":"632","article-title":"A fast fusion object determination method for multi-path video and three-dimensional GIS scene","volume":"49","author":"Li","year":"2020","journal-title":"Acta Geod. Cartogr. Sin."},{"key":"ref_30","doi-asserted-by":"crossref","first-page":"88133","DOI":"10.1109\/ACCESS.2020.2989157","article-title":"Fast SIFT feature matching algorithm based on geometric transformation","volume":"8","author":"Wang","year":"2020","journal-title":"IEEE Access"},{"key":"ref_31","first-page":"1395","article-title":"Research of image matching based on improved SURF algorithm","volume":"12","author":"Qi","year":"2014","journal-title":"TELKOMNIKA Indones. J. Electr. Eng."},{"key":"ref_32","doi-asserted-by":"crossref","first-page":"012151","DOI":"10.1088\/1742-6596\/1871\/1\/012151","article-title":"Improved ORB matching algorithm based on adaptive threshold","volume":"1871","author":"Li","year":"2021","journal-title":"J. Phys. Conf. Ser."},{"key":"ref_33","doi-asserted-by":"crossref","unstructured":"Zhang, J., Zhou, H., Niu, Y., Lv, J., Chen, J., and Cheng, Y. (2021). CNN and multi-feature extraction based denoising of CT images. Biomed. Signal Process. Control, 67.","DOI":"10.1016\/j.bspc.2021.102545"},{"key":"ref_34","doi-asserted-by":"crossref","unstructured":"Hong, Y., Li, D., Luo, S., Chen, X., Yang, Y., and Wang, M. (2022). An improved end-to-end multi-target tracking method based on transformer self-attention. Remote Sens., 14.","DOI":"10.3390\/rs14246354"},{"key":"ref_35","doi-asserted-by":"crossref","first-page":"1127","DOI":"10.1016\/j.procs.2023.10.624","article-title":"Deep feature-based RGB-D odometry using SuperPoint and SuperGlue","volume":"227","author":"Fujimoto","year":"2023","journal-title":"Procedia Comput. Sci."},{"key":"ref_36","unstructured":"Shen, X., Cai, Z., Yin, W., M\u00fcller, M., Li, Z., Wang, K., Chen, X.Z., and Wang, C. (2024). GIM: Learning generalizable image matcher from internet videos. arXiv."},{"key":"ref_37","doi-asserted-by":"crossref","first-page":"5703410","DOI":"10.1109\/TGRS.2023.3292372","article-title":"EVAA\u2014Exchange vanishing adversarial attack on LiDAR point clouds in autonomous vehicles","volume":"61","author":"Vishnu","year":"2023","journal-title":"IEEE Trans. Geosci. Remote Sens."},{"key":"ref_38","doi-asserted-by":"crossref","unstructured":"Chen, L., Zhu, Y., Papandreou, G., Schroff, F., and Adam, H. (2018, January 8\u201314). Encoder-decoder with atrous separable convolution for semantic image segmentation. Proceedings of the European Conference on Computer Vision (ECCV), Munich, Germany.","DOI":"10.1007\/978-3-030-01234-2_49"}],"container-title":["ISPRS International Journal of Geo-Information"],"original-title":[],"language":"en","link":[{"URL":"https:\/\/www.mdpi.com\/2220-9964\/14\/10\/398\/pdf","content-type":"unspecified","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2025,10,15]],"date-time":"2025-10-15T04:38:28Z","timestamp":1760503108000},"score":1,"resource":{"primary":{"URL":"https:\/\/www.mdpi.com\/2220-9964\/14\/10\/398"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2025,10,13]]},"references-count":38,"journal-issue":{"issue":"10","published-online":{"date-parts":[[2025,10]]}},"alternative-id":["ijgi14100398"],"URL":"https:\/\/doi.org\/10.3390\/ijgi14100398","relation":{},"ISSN":["2220-9964"],"issn-type":[{"value":"2220-9964","type":"electronic"}],"subject":[],"published":{"date-parts":[[2025,10,13]]}}}