{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2026,1,29]],"date-time":"2026-01-29T21:50:20Z","timestamp":1769723420403,"version":"3.49.0"},"reference-count":37,"publisher":"Springer Science and Business Media LLC","issue":"14","license":[{"start":{"date-parts":[[2022,5,7]],"date-time":"2022-05-07T00:00:00Z","timestamp":1651881600000},"content-version":"tdm","delay-in-days":0,"URL":"https:\/\/creativecommons.org\/licenses\/by\/4.0"},{"start":{"date-parts":[[2022,5,7]],"date-time":"2022-05-07T00:00:00Z","timestamp":1651881600000},"content-version":"vor","delay-in-days":0,"URL":"https:\/\/creativecommons.org\/licenses\/by\/4.0"}],"funder":[{"name":"Shenzhen Science and Technology Innovation Committee","award":["JCYJ20190813170803617"],"award-info":[{"award-number":["JCYJ20190813170803617"]}]}],"content-domain":{"domain":["link.springer.com"],"crossmark-restriction":false},"short-container-title":["J Supercomput"],"published-print":{"date-parts":[[2022,9]]},"abstract":"<jats:title>Abstract<\/jats:title><jats:p>Large-scale high-resolution three-dimensional (3D) maps play a vital role in the development of smart cities. In this work, a novel deep learning-based multi-view-stereo method is proposed for reconstructing the 3D maps in large-scale urban environments by exploiting a monocular camera. Compared with other existing works, the proposed method can perform 3D depth estimation more efficiently in terms of computational complexity and graphics processing unit memory usage. As a result, the proposed method can practically perform depth estimation for each pixel before generating 3D maps for even large-scale scenes. Extensive experiments on the well-known DTU dataset and real-life data collected on our campus confirm the good performance of the proposed method.<\/jats:p>","DOI":"10.1007\/s11227-022-04512-5","type":"journal-article","created":{"date-parts":[[2022,5,7]],"date-time":"2022-05-07T04:02:47Z","timestamp":1651896167000},"page":"16512-16528","update-policy":"https:\/\/doi.org\/10.1007\/springer_crossmark_policy","source":"Crossref","is-referenced-by-count":12,"title":["3D map reconstruction using a monocular camera for smart cities"],"prefix":"10.1007","volume":"78","author":[{"given":"Yuxi","family":"Hu","sequence":"first","affiliation":[],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Taimeng","family":"Fu","sequence":"additional","affiliation":[],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Guanchong","family":"Niu","sequence":"additional","affiliation":[],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Zixiao","family":"Liu","sequence":"additional","affiliation":[],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Man-On","family":"Pun","sequence":"additional","affiliation":[],"role":[{"role":"author","vocabulary":"crossref"}]}],"member":"297","published-online":{"date-parts":[[2022,5,7]]},"reference":[{"issue":"3","key":"4512_CR1","doi-asserted-by":"publisher","first-page":"494","DOI":"10.1109\/TSC.2014.2347293","volume":"8","author":"G Huang","year":"2014","unstructured":"Huang G, Ma Y, Liu X, Luo Y, Lu X, Blake MB (2014) Model-based automated navigation and composition of complex service mashups. IEEE Trans Serv Comput 8(3):494\u2013506","journal-title":"IEEE Trans Serv Comput"},{"issue":"1","key":"4512_CR2","doi-asserted-by":"publisher","first-page":"6","DOI":"10.1109\/TSC.2016.2587260","volume":"12","author":"G Huang","year":"2016","unstructured":"Huang G, Liu X, Ma Y, Lu X, Zhang Y, Xiong Y (2016) Programming situational mobile web applications with cloud-mobile convergence: an internetware-oriented approach. IEEE Trans Serv Comput 12(1):6\u201319","journal-title":"IEEE Trans Serv Comput"},{"issue":"11","key":"4512_CR3","first-page":"1","volume":"62","author":"X Chen","year":"2019","unstructured":"Chen X, Lin J, Ma Y, Lin B, Wang H, Huang G (2019) Self-adaptive resource allocation for cloud-based software services based on progressive qos prediction model. Science China Inf Sci 62(11):1\u20133","journal-title":"Science China Inf Sci"},{"key":"4512_CR4","doi-asserted-by":"publisher","first-page":"287","DOI":"10.1016\/j.future.2019.12.005","volume":"105","author":"X Chen","year":"2020","unstructured":"Chen X, Wang H, Ma Y, Zheng X, Guo L (2020) Self-adaptive resource allocation for cloud-based software services based on iterative qos pre- diction model. Futur Gener Comput Syst 105:287\u2013296","journal-title":"Futur Gener Comput Syst"},{"issue":"10","key":"4512_CR5","doi-asserted-by":"publisher","first-page":"2913","DOI":"10.1109\/TMC.2017.2651823","volume":"16","author":"G Huang","year":"2017","unstructured":"Huang G, Xu M, Lin FX, Liu Y, Ma Y, Pushp S, Liu X (2017) Shuffle- dog: Characterizing and adapting user-perceived latency of android apps. IEEE Trans Mob Comput 16(10):2913\u20132926","journal-title":"IEEE Trans Mob Comput"},{"issue":"4","key":"4512_CR6","doi-asserted-by":"publisher","first-page":"540","DOI":"10.1007\/s11704-015-4362-0","volume":"9","author":"X Chen","year":"2015","unstructured":"Chen X, Li A, Guo W, Huang G et al (2015) Runtime model based approach to iot application development. Front Comp Sci 9(4):540\u2013553","journal-title":"Front Comp Sci"},{"key":"4512_CR7","doi-asserted-by":"publisher","first-page":"1208","DOI":"10.1016\/j.ins.2020.10.001","volume":"546","author":"C-M Chen","year":"2021","unstructured":"Chen C-M, Chen L, Gan W, Qiu L, Ding W (2021) Discovering high utility-occupancy patterns from uncertain data. Inf Sci 546:1208\u20131229","journal-title":"Inf Sci"},{"key":"4512_CR8","unstructured":"Chen C-M, Huang Y, Wang K-H, Kumari S, Wu M-E (2020) A secure authenticated and key exchange scheme for fog computing. Enterpr Inf Syst, 1\u201316"},{"issue":"1","key":"4512_CR9","doi-asserted-by":"publisher","first-page":"1","DOI":"10.1007\/s11432-015-5499-z","volume":"57","author":"X Liu","year":"2014","unstructured":"Liu X, Huang G, Zhao Q, Mei H, Blake MB (2014) imashup: a mashup- based framework for service composition. Science China Inf Sci 57(1):1\u201320","journal-title":"Science China Inf Sci"},{"issue":"8","key":"4512_CR10","doi-asserted-by":"publisher","first-page":"5456","DOI":"10.1109\/TII.2019.2961237","volume":"16","author":"B Lin","year":"2019","unstructured":"Lin B, Huang Y, Zhang J, Hu J, Chen X, Li J (2019) Cost-driven off- loading for dnn-based applications over cloud, edge, and end devices. IEEE Trans Industr Inf 16(8):5456\u20135466","journal-title":"IEEE Trans Industr Inf"},{"issue":"2","key":"4512_CR11","doi-asserted-by":"publisher","first-page":"153","DOI":"10.1007\/s11263-016-0902-9","volume":"120","author":"H Aan\u00e6s","year":"2016","unstructured":"Aan\u00e6s H, Jensen RR, Vogiatzis G, Tola E, Dahl AB (2016) Large- scale data for multiple-view stereopsis. Int J Comput Vision 120(2):153\u2013168","journal-title":"Int J Comput Vision"},{"issue":"4","key":"4512_CR12","first-page":"1","volume":"36","author":"P-S Wang","year":"2017","unstructured":"Wang P-S, Liu Y, Guo Y-X, Sun C-Y, Tong X (2017) O-cnn: Octree- based convolutional neural networks for 3d shape analysis. ACM Trans- actions On Graphics (TOG) 36(4):1\u201311","journal-title":"ACM Trans- actions On Graphics (TOG)"},{"key":"4512_CR13","doi-asserted-by":"crossref","unstructured":"Pang, J., Sun, W., Ren, J.S., Yang, C., Yan, Q.: Cascade residual learn- ing: A two-stage convolutional neural network for stereo matching. In: Proceedings of the IEEE International Conference on Computer Vision Workshops, pp. 887\u2013895 (2017)","DOI":"10.1109\/ICCVW.2017.108"},{"key":"4512_CR14","doi-asserted-by":"crossref","unstructured":"Wu Z, Wu X, Zhang X, Wang S, Ju L (2019) Semantic stereo match- ing with pyramid cost volumes. In: Proceedings of the IEEE\/CVF international conference on computer vision, pp 7484\u20137493","DOI":"10.1109\/ICCV.2019.00758"},{"key":"4512_CR15","doi-asserted-by":"crossref","unstructured":"Liang Z, Feng Y, Guo Y, Liu H, Chen W, Qiao L, Zhou L, Zhang J (2018) Learning for disparity estimation through feature constancy. In: Proceedings of the IEEE conference on computer vision and pattern recognition, pp 2811\u20132820","DOI":"10.1109\/CVPR.2018.00297"},{"key":"4512_CR16","unstructured":"Kar A, H\u00a8ane C, Malik J (2017) Learning a multi-view stereo machine. arXiv preprint arXiv:1708.05375"},{"issue":"5","key":"4512_CR17","doi-asserted-by":"publisher","first-page":"903","DOI":"10.1007\/s00138-011-0346-8","volume":"23","author":"E Tola","year":"2012","unstructured":"Tola E, Strecha C, Fua P (2012) Efficient large-scale multi-view stereo for ultra high-resolution image sets. Mach Vis Appl 23(5):903\u2013920","journal-title":"Mach Vis Appl"},{"key":"4512_CR18","doi-asserted-by":"crossref","unstructured":"Yao Y, Luo Z, Li S, Shen T, Fang T, Quan L (2019) Recurrent mvsnet for high-resolution multi-view stereo depth inference. In: Proceedings of the IEEE\/CVF conference on computer vision and pattern recognition, pp 5525\u20135534","DOI":"10.1109\/CVPR.2019.00567"},{"key":"4512_CR19","doi-asserted-by":"crossref","unstructured":"Riegler G, Osman Ulusoy A, Geiger A (2017) Octnet: Learning deep 3d rep- resentations at high resolutions. In: Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition, pp 3577\u20133586","DOI":"10.1109\/CVPR.2017.701"},{"issue":"3","key":"4512_CR20","doi-asserted-by":"publisher","first-page":"418","DOI":"10.1109\/TPAMI.2005.44","volume":"27","author":"M Lhuillier","year":"2005","unstructured":"Lhuillier M, Quan L (2005) A quasi-dense approach to surface reconstruc- tion from uncalibrated images. IEEE Trans Pattern Anal Mach Intell 27(3):418\u2013433","journal-title":"IEEE Trans Pattern Anal Mach Intell"},{"issue":"8","key":"4512_CR21","doi-asserted-by":"publisher","first-page":"1362","DOI":"10.1109\/TPAMI.2009.161","volume":"32","author":"Y Furukawa","year":"2009","unstructured":"Furukawa Y, Ponce J (2009) Accurate, dense, and robust multiview stereopsis. IEEE Trans Pattern Anal Mach Intell 32(8):1362\u20131376","journal-title":"IEEE Trans Pattern Anal Mach Intell"},{"key":"4512_CR22","doi-asserted-by":"crossref","unstructured":"Sinha SN, Mordohai P, Pollefeys M (2007) Multi-view stereo via graph cuts on the dual of an adaptive tetrahedral mesh. In: 2007 IEEE 11th international conference on computer vision, pp 1\u20138. IEEE","DOI":"10.1109\/ICCV.2007.4408997"},{"key":"4512_CR23","doi-asserted-by":"crossref","unstructured":"Furukawa Y, Ponce J (2006) Carved visual hulls for image-based modeling. In: European conference on computer vision, pp 564\u2013577 Springer","DOI":"10.1007\/11744023_44"},{"key":"4512_CR24","unstructured":"Galliani, S., Lasinger, K., Schindler, K.: Gipuma: Massively parallel multi- view stereo reconstruction"},{"key":"4512_CR25","doi-asserted-by":"crossref","unstructured":"Schonberger JL, Frahm, J-M (2016) Structure-from-motion revisited. In: Proceedings of the IEEE conference on computer vision and pattern recognition, pp 4104\u20134113","DOI":"10.1109\/CVPR.2016.445"},{"key":"4512_CR26","doi-asserted-by":"crossref","unstructured":"Huang Y, Wang L (2019) Acmm: Aligned cross-modal memory for few- shot image and sentence matching. In: Proceedings of the IEEE\/CVF international conference on computer vision, pp 5774\u20135783","DOI":"10.1109\/ICCV.2019.00587"},{"key":"4512_CR27","doi-asserted-by":"crossref","unstructured":"Ji M, Gall J, Zheng H, Liu Y, Fang L (2017) Surfacenet: an end-to-end 3d neural network for multiview stereopsis. In: Proceedings of the IEEE international conference on computer vision, pp 2307\u20132315","DOI":"10.1109\/ICCV.2017.253"},{"key":"4512_CR28","doi-asserted-by":"crossref","unstructured":"Huang P-H, Matzen K, Kopf J, Ahuja N, Huang J-B (2018) Deepmvs: Learning multi-view stereopsis. In: Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition, pp 2821\u20132830","DOI":"10.1109\/CVPR.2018.00298"},{"key":"4512_CR29","doi-asserted-by":"crossref","unstructured":"Kendall A, Martirosyan H, Dasgupta S, Henry P, Kennedy R, Bachrach A, Bry A (2017) End-to-end learning of geometry and context for deep stereo regression. In: Proceedings of the IEEE international conference on computer vision, pp 66\u201375","DOI":"10.1109\/ICCV.2017.17"},{"key":"4512_CR30","doi-asserted-by":"crossref","unstructured":"Chen R, Han S, Xu J, Su H (2019) Point-based multi-view stereo network. In: Proceedings of the IEEE\/CVF international conference on computer vision, pp 1538\u20131547","DOI":"10.1109\/ICCV.2019.00162"},{"key":"4512_CR31","doi-asserted-by":"crossref","unstructured":"Lin T-Y, Doll\u00b4ar P, Girshick R, He K, Hariharan B, Belongie S (2017) Feature pyramid networks for object detection. In: Proceedings of the IEEE conference on computer vision and pattern recognition, pp 2117\u2013 2125","DOI":"10.1109\/CVPR.2017.106"},{"key":"4512_CR32","doi-asserted-by":"crossref","unstructured":"Yao Y, Luo Z, Li S., Fang, T., Quan L (2018) Mvsnet: Depth infer- ence for unstructured multi-view stereo. In: Proceedings of the European conference on computer vision (ECCV), pp 767\u2013783","DOI":"10.1007\/978-3-030-01237-3_47"},{"key":"4512_CR33","doi-asserted-by":"crossref","unstructured":"Song X, Zhao X, Hu H, Fang L (2018) Edgestereo: a context integrated residual pyramid network for stereo matching. In: Asian Conference on Computer Vision, pp 20\u201335. Springer","DOI":"10.1007\/978-3-030-20873-8_2"},{"key":"4512_CR34","doi-asserted-by":"crossref","unstructured":"Duggal S, Wang S, Ma W-C, Hu R, Urtasun R (2019) Deeppruner: learning efficient stereo matching via differentiable patchmatch. In: Proceedings of the IEEE\/CVF international conference on computer vision, pp 4384\u20134393","DOI":"10.1109\/ICCV.2019.00448"},{"issue":"7","key":"4512_CR35","doi-asserted-by":"publisher","first-page":"787","DOI":"10.1109\/TPAMI.2003.1206509","volume":"25","author":"J Sun","year":"2003","unstructured":"Sun J, Zheng N-N, Shum H-Y (2003) Stereo matching using belief propa- gation. IEEE Trans Pattern Anal Mach Intell 25(7):787\u2013800","journal-title":"IEEE Trans Pattern Anal Mach Intell"},{"key":"4512_CR36","doi-asserted-by":"crossref","unstructured":"Sch\u00a8onberger JL, Zheng E, Frahm J-M, Pollefeys M (2016) Pixelwise view selection for unstructured multi-view stereo. In: European conference on computer vision, pp 501\u2013518. Springer","DOI":"10.1007\/978-3-319-46487-9_31"},{"key":"4512_CR37","doi-asserted-by":"crossref","unstructured":"Gu X, Fan Z, Zhu S, Dai Z, Tan F, Tan P (2020) Cascade cost volume for high-resolution multi-view stereo and stereo matching. In: Proceedings of the IEEE\/CVF conference on computer vision and pattern recognition, pp 2495\u20132504","DOI":"10.1109\/CVPR42600.2020.00257"}],"container-title":["The Journal of Supercomputing"],"original-title":[],"language":"en","link":[{"URL":"https:\/\/link.springer.com\/content\/pdf\/10.1007\/s11227-022-04512-5.pdf","content-type":"application\/pdf","content-version":"vor","intended-application":"text-mining"},{"URL":"https:\/\/link.springer.com\/article\/10.1007\/s11227-022-04512-5\/fulltext.html","content-type":"text\/html","content-version":"vor","intended-application":"text-mining"},{"URL":"https:\/\/link.springer.com\/content\/pdf\/10.1007\/s11227-022-04512-5.pdf","content-type":"application\/pdf","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2022,9,8]],"date-time":"2022-09-08T16:48:35Z","timestamp":1662655715000},"score":1,"resource":{"primary":{"URL":"https:\/\/link.springer.com\/10.1007\/s11227-022-04512-5"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2022,5,7]]},"references-count":37,"journal-issue":{"issue":"14","published-print":{"date-parts":[[2022,9]]}},"alternative-id":["4512"],"URL":"https:\/\/doi.org\/10.1007\/s11227-022-04512-5","relation":{},"ISSN":["0920-8542","1573-0484"],"issn-type":[{"value":"0920-8542","type":"print"},{"value":"1573-0484","type":"electronic"}],"subject":[],"published":{"date-parts":[[2022,5,7]]},"assertion":[{"value":"29 March 2022","order":1,"name":"accepted","label":"Accepted","group":{"name":"ArticleHistory","label":"Article History"}},{"value":"7 May 2022","order":2,"name":"first_online","label":"First Online","group":{"name":"ArticleHistory","label":"Article History"}}]}}