{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2025,6,5]],"date-time":"2025-06-05T10:51:56Z","timestamp":1749120716666,"version":"3.37.3"},"reference-count":75,"publisher":"Springer Science and Business Media LLC","issue":"11","license":[{"start":{"date-parts":[[2022,9,13]],"date-time":"2022-09-13T00:00:00Z","timestamp":1663027200000},"content-version":"tdm","delay-in-days":0,"URL":"https:\/\/creativecommons.org\/licenses\/by\/4.0"},{"start":{"date-parts":[[2022,9,13]],"date-time":"2022-09-13T00:00:00Z","timestamp":1663027200000},"content-version":"vor","delay-in-days":0,"URL":"https:\/\/creativecommons.org\/licenses\/by\/4.0"}],"funder":[{"DOI":"10.13039\/501100000266","name":"Engineering and Physical Sciences Research Council","doi-asserted-by":"publisher","award":["EP\/N509772\/1","EP\/P022529\/1"],"award-info":[{"award-number":["EP\/N509772\/1","EP\/P022529\/1"]}],"id":[{"id":"10.13039\/501100000266","id-type":"DOI","asserted-by":"publisher"}]}],"content-domain":{"domain":["link.springer.com"],"crossmark-restriction":false},"short-container-title":["Int J Comput Vis"],"published-print":{"date-parts":[[2022,11]]},"abstract":"<jats:title>Abstract<\/jats:title><jats:p>Multi-view stereo remains a popular choice when recovering 3D geometry, despite performance varying dramatically according to the scene content. Moreover, typical pinhole camera assumptions fail in the presence of shallow depth of field inherent to macro-scale scenes; limiting application to larger scenes with diffuse reflectance. However, the presence of defocus blur can itself be considered a useful reconstruction cue, particularly in the presence of view-dependent materials. With this in mind, we explore the complimentary nature of stereo and defocus cues in the context of multi-view 3D reconstruction; and propose a complete pipeline for scene modelling from a finite aperature camera that encompasses image formation, camera calibration and reconstruction stages. As part of our evaluation, an ablation study reveals how each cue contributes to the higher performance observed over a range of complex materials and geometries. Though of lesser concern with large apertures, the effects of image noise are also considered. By introducing pre-trained deep feature extraction into our cost function, we show a step improvement over per-pixel comparisons; as well as verify the cross-domain applicability of networks using largely in-focus training data applied to defocused images. Finally, we compare to a number of modern multi-view stereo methods, and demonstrate how the use of both cues leads to a significant increase in performance across several synthetic and real datasets.<\/jats:p>","DOI":"10.1007\/s11263-022-01658-w","type":"journal-article","created":{"date-parts":[[2022,9,13]],"date-time":"2022-09-13T09:03:28Z","timestamp":1663059808000},"page":"2858-2884","update-policy":"https:\/\/doi.org\/10.1007\/springer_crossmark_policy","source":"Crossref","is-referenced-by-count":1,"title":["Finite Aperture Stereo"],"prefix":"10.1007","volume":"130","author":[{"ORCID":"https:\/\/orcid.org\/0000-0003-4638-3214","authenticated-orcid":false,"given":"Matthew","family":"Bailey","sequence":"first","affiliation":[],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Adrian","family":"Hilton","sequence":"additional","affiliation":[],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Jean-Yves","family":"Guillemaut","sequence":"additional","affiliation":[],"role":[{"role":"author","vocabulary":"crossref"}]}],"member":"297","published-online":{"date-parts":[[2022,9,13]]},"reference":[{"key":"1658_CR1","doi-asserted-by":"crossref","unstructured":"Acharyya, A., Hudson, D., Chen, K.W., Feng, T., Kan, C.-Y., & Nguyen, T. (2016). Depth estimation from focus and disparity. In IEEE international conference on image processing (ICIP) (pp. 3444\u20133448).","DOI":"10.1109\/ICIP.2016.7532999"},{"key":"1658_CR2","doi-asserted-by":"crossref","unstructured":"Anwar, S., Hayder, Z., & Porikli, F. (2021). Deblur and deep depth from single defocus image. Machine Vision and Applications32(1).","DOI":"10.1007\/s00138-020-01162-6"},{"key":"1658_CR3","doi-asserted-by":"publisher","unstructured":"Bailey, M., & Guillemaut, J.-Y. (2020). A novel depth from defocus framework based on a thick lens camera model. In 2020 international conference on 3D vision (3DV) (pp. 1206\u20131215). https:\/\/doi.org\/10.1109\/3DV50981.2020.00131","DOI":"10.1109\/3DV50981.2020.00131"},{"key":"1658_CR4","doi-asserted-by":"publisher","unstructured":"Bailey, M., Hilton, A., & Guillemaut, J.-Y. (2021). Finite aperture stereo: 3D reconstruction of macro-scale scenes. In 2021 IEEE\/CVF International Conference on Computer Vision Workshops (ICCVW) (pp. 2474\u20132484). https:\/\/doi.org\/10.1109\/ICCVW54120.2021.00280","DOI":"10.1109\/ICCVW54120.2021.00280"},{"issue":"6","key":"1658_CR5","doi-asserted-by":"publisher","first-page":"1041","DOI":"10.1109\/TPAMI.2014.14","volume":"36","author":"R Ben-Ari","year":"2014","unstructured":"Ben-Ari, R. (2014). a unified approach for registration and depth in depth from defocus. IEEE Transactions on Pattern Analysis and Machine Intelligence, 36(6), 1041\u20131055.","journal-title":"IEEE Transactions on Pattern Analysis and Machine Intelligence"},{"issue":"2","key":"1658_CR6","doi-asserted-by":"publisher","first-page":"167","DOI":"10.1007\/s11263-011-0476-5","volume":"97","author":"A Bhavsar","year":"2012","unstructured":"Bhavsar, A., & Rajagopalan, A. (2012). Towards unrestrained depth inference with coherent occlusion filling. International Journal of Computer Vision, 97(2), 167\u2013190.","journal-title":"International Journal of Computer Vision"},{"issue":"11","key":"1658_CR7","doi-asserted-by":"publisher","first-page":"1222","DOI":"10.1109\/34.969114","volume":"23","author":"Y Boykov","year":"2001","unstructured":"Boykov, Y., Veksler, O., & Zabih, R. (2001). Fast approximate energy minimization via graph cuts. IEEE Transactions on Pattern Analysis and Machine Intelligence, 23(11), 1222\u20131239.","journal-title":"IEEE Transactions on Pattern Analysis and Machine Intelligence"},{"key":"1658_CR8","doi-asserted-by":"crossref","unstructured":"Bradley, D., Boubekeur, T., Heidrich, W. (2008) Accurate multi-view reconstruction using robust binocular stereo and surface meshing. In IEEE conference on computer vision and pattern recognition, pp. 1\u20138.","DOI":"10.1109\/CVPR.2008.4587792"},{"key":"1658_CR9","doi-asserted-by":"crossref","unstructured":"Carvalho, M., Le\u00a0Saux, B., Trouv\u00e9-Peloux, P., Almansa, A., Champagnat, F. (2019) Deep depth from defocus: How can defocus blur improve 3D estimation using dense neural networks? In Computer Vision \u2013 ECCV 2018 Workshops (pp. 307\u2013323). Springer, Cham.","DOI":"10.1007\/978-3-030-11009-3_18"},{"key":"1658_CR10","doi-asserted-by":"crossref","unstructured":"Chakrabarti, A., Zickler, T. (2012). Depth and deblurring from a spectrally-varying depth-of-field. In Computer vision - ECCV. Lecture Notes in Computer Science (pp. 648\u2013661). Springer, Berlin, Heidelberg.","DOI":"10.1007\/978-3-642-33715-4_47"},{"key":"1658_CR11","unstructured":"Chen, Z., Guo, X., Li, S., Cao, X., & Yu, J. (2017). A Learning-based framework for hybrid depth-from-defocus and stereo matching. arXiv e-prints arXiv:1708.00583 [cs.CV]"},{"key":"1658_CR12","doi-asserted-by":"crossref","unstructured":"Chen, R., Han, S., Xu, J., Su, H. (2019) Point-based multi-view stereo network. CoRR abs\/1908.04422. arXiv:1908.04422.","DOI":"10.1109\/ICCV.2019.00162"},{"key":"1658_CR13","doi-asserted-by":"crossref","unstructured":"Chen, L. Y., Shuochen, S., & S., Matsushita, S., KunZhou, S., & Lin, S. (2016). Bayesian depth-from-defocus with shading constraints. IEEE Transactions on Image Processing,25(2), 589\u2013600.","DOI":"10.1109\/TIP.2015.2507403"},{"key":"1658_CR14","doi-asserted-by":"crossref","unstructured":"Chen, C.-H., Zhou, H., & Ahonen, T. (2015). Blur-Aware Disparity Estimation from Defocus stereo images. In 2015 IEEE international conference on computer vision (ICCV) (Vol. 2015, pp. 855\u2013863).","DOI":"10.1109\/ICCV.2015.104"},{"key":"1658_CR15","doi-asserted-by":"crossref","unstructured":"Choy, C.B., Xu, D., Gwak, J., Chen, K., & Savarese, S. (2016). 3D-R2N2: A unified approach for single and multi-view 3D object reconstruction. In Proceedings of the European conference on computer vision (ECCV).","DOI":"10.1007\/978-3-319-46484-8_38"},{"key":"1658_CR16","doi-asserted-by":"publisher","unstructured":"Delaunoy, A., Pollefeys, M. (2014). Photometric bundle adjustment for dense multi-view 3d modeling. In 2014 IEEE conference on computer vision and pattern recognition (pp. 1486\u20131493). https:\/\/doi.org\/10.1109\/CVPR.2014.193.","DOI":"10.1109\/CVPR.2014.193"},{"key":"1658_CR17","doi-asserted-by":"publisher","unstructured":"Emerson, D.R., & Christopher, L.A. (2019). 3-D scene reconstruction using depth from defocus and deep learning. In 2019 IEEE applied imagery pattern recognition workshop (AIPR) (pp. 1\u20138). https:\/\/doi.org\/10.1109\/AIPR47015.2019.9174568","DOI":"10.1109\/AIPR47015.2019.9174568"},{"key":"1658_CR18","doi-asserted-by":"crossref","unstructured":"Favaro, P. (2007). Shape from focus and defocus: Convexity, quasiconvexity and defocus-invariant textures. In IEEE 11th international conference on computer vision (ICCV) (pp. 1\u20137).","DOI":"10.1109\/ICCV.2007.4409024"},{"key":"1658_CR19","doi-asserted-by":"crossref","unstructured":"Favaro, P. (2010). Recovering thin structures via nonlocal-means regularization with application to depth from defocus. In IEEE conference on computer vision and pattern recognition (CVPR) (pp. 1133\u20131140).","DOI":"10.1109\/CVPR.2010.5540089"},{"issue":"3","key":"1658_CR20","doi-asserted-by":"publisher","first-page":"406","DOI":"10.1109\/TPAMI.2005.43","volume":"27","author":"P Favaro","year":"2005","unstructured":"Favaro, P., & Soatto, S. (2005). A geometric approach to shape from defocus. IEEE Transactions on Pattern Analysis and Machine Intelligence, 27(3), 406\u2013417.","journal-title":"IEEE Transactions on Pattern Analysis and Machine Intelligence"},{"issue":"3","key":"1658_CR21","doi-asserted-by":"publisher","first-page":"518","DOI":"10.1109\/TPAMI.2007.1175","volume":"30","author":"P Favaro","year":"2008","unstructured":"Favaro, P., Soatto, S., Burger, M., & Osher, S. J. (2008). Shape from defocus via diffusion. IEEE Transactions on Pattern Analysis and Machine Intelligence, 30(3), 518\u2013531.","journal-title":"IEEE Transactions on Pattern Analysis and Machine Intelligence"},{"issue":"8","key":"1658_CR22","doi-asserted-by":"publisher","first-page":"1362","DOI":"10.1109\/TPAMI.2009.161","volume":"32","author":"Y Furukawa","year":"2010","unstructured":"Furukawa, Y., & Ponce, J. (2010). Accurate, dense, and robust multiview stereopsis. IEEE Transactions On Pattern Analysis And Machine Intelligence, 32(8), 1362\u20131376.","journal-title":"IEEE Transactions On Pattern Analysis And Machine Intelligence"},{"key":"1658_CR23","first-page":"26","volume-title":"Informatik 2007 - Informatik Trifft Logistik -","author":"I Ghe\u0163a","year":"2007","unstructured":"Ghe\u0163a, I., Frese, C., Heizmann, M., & Beyerer, J. (2007). A new approach for estimating depth by fusing stereo and defocus information. In R. Koschke, O. Herzog, K.-H. R\u00f6diger, & M. Ronthaler (Eds.), Informatik 2007 - Informatik Trifft Logistik - (Vol. 1, pp. 26\u201331). Bonn: Gesellschaft f\u00fcr Informatik e. V."},{"key":"1658_CR24","doi-asserted-by":"crossref","unstructured":"Gu, X., Fan, Z., Dai, Z., Zhu, S., Tan, F., & Tan, P. (2019). Cascade cost volume for high-resolution multi-view stereo and stereo matching (pp. 1912\u201306378).","DOI":"10.1109\/CVPR42600.2020.00257"},{"key":"1658_CR25","volume-title":"Multiple view geometry in computer vision","author":"R Hartley","year":"2000","unstructured":"Hartley, R. (2000). Multiple view geometry in computer vision. Cambridge: Cambridge University Press."},{"issue":"1","key":"1658_CR26","doi-asserted-by":"publisher","first-page":"82","DOI":"10.1007\/s11263-008-0164-2","volume":"81","author":"S Hasinoff","year":"2009","unstructured":"Hasinoff, S., & Kutulakos, K. (2009). Confocal stereo. International Journal of Computer Vision, 81(1), 82\u2013104.","journal-title":"International Journal of Computer Vision"},{"key":"1658_CR27","doi-asserted-by":"publisher","unstructured":"Hornung, A., Kobbelt, L. (2006). Hierarchical volumetric multi-view stereo reconstruction of manifold surfaces based on dual graph embedding. In IEEE computer society conference on computer vision and pattern recognition (CVPR) (Vol. 1, pp. 503\u2013510). https:\/\/doi.org\/10.1109\/CVPR.2006.135","DOI":"10.1109\/CVPR.2006.135"},{"key":"1658_CR28","doi-asserted-by":"crossref","unstructured":"Huang, P., Matzen, K., Kopf, J., Ahuja, N., & Huang, J. (2018). DeepMVS: Learning multi-view stereopsis. CoRR arXiv:1804.00650.","DOI":"10.1109\/CVPR.2018.00298"},{"key":"1658_CR29","doi-asserted-by":"crossref","unstructured":"Ji, M., Gall, J., Zheng, H., Liu, Y., Fang, L. (2017). Surfacenet: An end-to-end 3d neural network for multiview stereopsis. In Proceedings of the IEEE international conference on computer vision (ICCV) (pp. 2307\u20132315).","DOI":"10.1109\/ICCV.2017.253"},{"key":"1658_CR30","unstructured":"Kar, A., H\u00e4ne, C., Malik, J. (2017) Learning a multi-view stereo machine."},{"key":"1658_CR31","doi-asserted-by":"publisher","unstructured":"Kashiwagi, M., Mishima, N., Kozakaya, T., & Hiura, S. (2019). Deep depth from aberration map. In 2019 IEEE\/CVF international conference on computer vision (ICCV) (pp. 4069\u20134078). https:\/\/doi.org\/10.1109\/ICCV.2019.00417","DOI":"10.1109\/ICCV.2019.00417"},{"issue":"3","key":"1658_CR32","doi-asserted-by":"publisher","first-page":"1","DOI":"10.1145\/2487228.2487237","volume":"32","author":"M Kazhdan","year":"2013","unstructured":"Kazhdan, M., & Hoppe, H. (2013). Screened poisson surface reconstruction. ACM Transactions on Graphics, 32(3), 1\u201313.","journal-title":"ACM Transactions on Graphics"},{"key":"1658_CR33","doi-asserted-by":"crossref","unstructured":"Kingslake, R. (1992). Optics in photography. SPIE Press monograph; PM06. SPIE, Bellingham, Wash. (1000 20th St. Bellingham WA 98225-6705 USA)","DOI":"10.1117\/3.43160"},{"key":"1658_CR34","doi-asserted-by":"crossref","unstructured":"Knapitsch, A., Park, J., Zhou, Q.-Y., & Koltun, V. (2017). Tanks and temples: Benchmarking large-scale scene reconstruction. ACM Transactions on Graphics36(4).","DOI":"10.1145\/3072959.3073599"},{"key":"1658_CR35","doi-asserted-by":"publisher","unstructured":"Kuhn, A., Sormann, C., Rossi, M., Erdler, O., & Fraundorfer, F. (2020). DeepC-MVS: deep confidence prediction for multi-view stereo reconstruction. https:\/\/doi.org\/10.1109\/3DV50981.2020.00050","DOI":"10.1109\/3DV50981.2020.00050"},{"key":"1658_CR36","doi-asserted-by":"crossref","unstructured":"Li, F., Sun, J., Wang, J., & Yu, J. (2010). Dual-focus stereo imaging. Journal Of Electronic Imaging19(4).","DOI":"10.1117\/1.3500802"},{"key":"1658_CR37","doi-asserted-by":"crossref","unstructured":"Li, Z., Wang, K., Zuo, W., Meng, D., & Zhang, L. (2016). Detail-preserving and content-aware variational multi-view stereo reconstruction. IEEE Transactions on Image Processing25(2).","DOI":"10.1109\/TIP.2015.2507400"},{"issue":"1","key":"1658_CR38","doi-asserted-by":"publisher","first-page":"72","DOI":"10.1109\/TPAMI.2008.270","volume":"32","author":"G Li","year":"2010","unstructured":"Li, G., & Zucker, S. W. (2010). Differential geometric inference in surface stereo. IEEE Transactions on Pattern Analysis and Machine Intelligence, 32(1), 72\u201386.","journal-title":"IEEE Transactions on Pattern Analysis and Machine Intelligence"},{"key":"1658_CR39","doi-asserted-by":"publisher","unstructured":"Lin, H., Chen, C., Kang, S.B., & Yu, J. (2015). Depth recovery from light field using focal stack symmetry. In 2015 IEEE international conference on computer vision (ICCV) (pp. 3451\u20133459). https:\/\/doi.org\/10.1109\/ICCV.2015.394","DOI":"10.1109\/ICCV.2015.394"},{"key":"1658_CR40","doi-asserted-by":"crossref","unstructured":"Lin, X., Suo, J., Cao, X., & Dai, Q. (2013). Iterative feedback estimation of depth and radiance from defocused images. In Lecture Notes in Computer Science 11th Asian conference on computer vision (ACCV) (Vol. 7727, pp. 95\u2013109). Berlin, Heidelberg: Springer.","DOI":"10.1007\/978-3-642-37447-0_8"},{"issue":"3","key":"1658_CR41","doi-asserted-by":"publisher","first-page":"407","DOI":"10.1109\/TVCG.2009.88","volume":"16","author":"Y Liu","year":"2010","unstructured":"Liu, Y., Dai, Q., & Xu, W. (2010). A point-cloud-based multiview stereo algorithm for free-viewpoint video. IEEE Transactions on Visualization and Computer Graphics, 16(3), 407\u2013418.","journal-title":"IEEE Transactions on Visualization and Computer Graphics"},{"key":"1658_CR42","doi-asserted-by":"publisher","unstructured":"Logothetis, F., Mecca, R., & Cipolla, R. (2019). A differential volumetric approach to multi-view photometric stereo. In IEEE\/CVF international conference on computer vision (ICCV) (pp. 1052\u20131061). https:\/\/doi.org\/10.1109\/ICCV.2019.00114","DOI":"10.1109\/ICCV.2019.00114"},{"key":"1658_CR43","doi-asserted-by":"publisher","unstructured":"Luo, K., Guan, T., Ju, L., Huang, H., Luo, Y. (2019). P-MVSNet: Learning Patch-Wise Matching Confidence Aggregation for Multi-View Stereo. In 2019 IEEE\/CVF international conference on computer vision (ICCV) (pp. 10451\u201310460). https:\/\/doi.org\/10.1109\/ICCV.2019.01055","DOI":"10.1109\/ICCV.2019.01055"},{"key":"1658_CR44","doi-asserted-by":"crossref","unstructured":"Mannan, F., Langer, M.S. (2015). Optimal camera parameters for depth from defocus. In International conference on 3D vision (pp. 326\u2013334).","DOI":"10.1109\/3DV.2015.44"},{"key":"1658_CR45","doi-asserted-by":"crossref","unstructured":"Martinello, M., Wajs, A., Quan, S., Lee, H., Lim, C., Woo, T., Lee, W., Kim, S.-S., & Lee, D. (2015) Dual aperture photography: Image and depth from a mobile camera. In IEEE international conference on computational photography (ICCP) (pp. 1\u201310).","DOI":"10.1109\/ICCPHOT.2015.7168366"},{"key":"1658_CR46","doi-asserted-by":"publisher","first-page":"405","DOI":"10.1007\/978-3-030-58452-8_24","volume-title":"Computer Vision - ECCV 2020","author":"B Mildenhall","year":"2020","unstructured":"Mildenhall, B., Srinivasan, P. P., Tancik, M., Barron, J. T., Ramamoorthi, R., & Ng, R. (2020). Nerf: Representing scenes as neural radiance fields for view synthesis. In A. Vedaldi, H. Bischof, T. Brox, & J.-M. Frahm (Eds.), Computer Vision - ECCV 2020 (pp. 405\u2013421). Cham: Springer."},{"issue":"12","key":"1658_CR47","doi-asserted-by":"publisher","first-page":"5369","DOI":"10.1109\/TIP.2015.2479469","volume":"24","author":"M Moeller","year":"2015","unstructured":"Moeller, M., Benning, M., Schonlieb, C., & Cremers, D. (2015). Variational depth from focus reconstruction. IEEE Transactions on Image Processing, 24(12), 5369\u20135378.","journal-title":"IEEE Transactions on Image Processing"},{"key":"1658_CR48","doi-asserted-by":"crossref","unstructured":"Namboodiri, V.P., Chaudhuri, S., & Hadap, S. (2008). Regularized depth from defocus. In 15th IEEE international conference on image processing (ICIP) (pp. 1520\u20131523).","DOI":"10.1109\/ICIP.2008.4712056"},{"issue":"8","key":"1658_CR49","doi-asserted-by":"publisher","first-page":"824","DOI":"10.1109\/34.308479","volume":"16","author":"SK Nayar","year":"1994","unstructured":"Nayar, S. K., & Nakagawa, Y. (1994). Shape from focus. IEEE Transactions on Pattern Analysis and Machine Intelligence, 16(8), 824\u2013831. https:\/\/doi.org\/10.1109\/34.308479.","journal-title":"IEEE Transactions on Pattern Analysis and Machine Intelligence"},{"key":"1658_CR50","doi-asserted-by":"crossref","unstructured":"Olsson, C., Ul\u00e9n, J., Boykov, Y. (2013). In defense of 3D-label stereo. In IEEE Conference on Computer Vision and Pattern Recognition (pp. 1730\u20131737).","DOI":"10.1109\/CVPR.2013.226"},{"key":"1658_CR51","doi-asserted-by":"publisher","unstructured":"Paramonov, V., Panchenko, I., Bucha, V., Drogolyub, A., & Zagoruyko, S. (2016). Depth camera based on color-coded aperture. In 2016 IEEE conference on computer vision and pattern recognition workshops (CVPRW) (pp. 910\u2013918).https:\/\/doi.org\/10.1109\/CVPRW.2016.118","DOI":"10.1109\/CVPRW.2016.118"},{"issue":"4","key":"1658_CR52","doi-asserted-by":"publisher","first-page":"523","DOI":"10.1109\/TPAMI.1987.4767940","volume":"9","author":"AP Pentland","year":"1987","unstructured":"Pentland, A. P. (1987). A new sense for depth of field. IEEE Transactions on Pattern Analysis and Machine Intelligence PAMI, 9(4), 523\u2013531.","journal-title":"IEEE Transactions on Pattern Analysis and Machine Intelligence PAMI"},{"key":"1658_CR53","doi-asserted-by":"publisher","first-page":"114","DOI":"10.1016\/j.imavis.2016.08.011","volume":"57","author":"N Persch","year":"2017","unstructured":"Persch, N., Schroers, C., Setzer, S., & Weickert, J. (2017). Physically inspired depth-from-defocus. Image and Vision Computing, 57, 114\u2013129.","journal-title":"Image and Vision Computing"},{"key":"1658_CR54","doi-asserted-by":"crossref","unstructured":"Rajagopalan, A.N., Chaudhuri, S., & Mudenagudi, U. (2004). Depth estimation and image restoration using defocused stereo pairs. IEEE Transactions on Pattern Analysis and Machine Intelligence26(11).","DOI":"10.1109\/TPAMI.2004.102"},{"key":"1658_CR55","volume-title":"GrabCut: Interactive foreground extraction using iterated graph cuts","author":"C Rother","year":"2004","unstructured":"Rother, C., Kolmogorov, V., & Blake, A. (2004). GrabCut: Interactive foreground extraction using iterated graph cuts. New York, NY, USA: Association for Computing Machinery."},{"key":"1658_CR56","doi-asserted-by":"publisher","DOI":"10.1088\/978-0-7503-1242-4ch1","author":"A Rowlands","year":"2017","unstructured":"Rowlands, A. (2017). Fundamental optical formulae. Physics of Digital Photography. https:\/\/doi.org\/10.1088\/978-0-7503-1242-4ch1.","journal-title":"Physics of Digital Photography"},{"key":"1658_CR57","doi-asserted-by":"publisher","first-page":"501","DOI":"10.1007\/978-3-319-46487-9_31","volume-title":"Computer Vision - ECCV 2016","author":"JL Sch\u00f6nberger","year":"2016","unstructured":"Sch\u00f6nberger, J. L., Zheng, E., Frahm, J.-M., & Pollefeys, M. (2016). Pixelwise view selection for unstructured multi-view stereo. In B. Leibe, J. Matas, N. Sebe, & M. Welling (Eds.), Computer Vision - ECCV 2016 (pp. 501\u2013518). Cham: Springer."},{"key":"1658_CR58","doi-asserted-by":"crossref","unstructured":"Song, G., & Lee, K.M. (2018) Depth estimation network for dual defocused images with different depth-of-field. In 25th IEEE international conference on image processing (ICIP) (pp. 1563\u20131567).","DOI":"10.1109\/ICIP.2018.8451201"},{"issue":"6","key":"1658_CR59","doi-asserted-by":"publisher","first-page":"1068","DOI":"10.1109\/TPAMI.2007.70844","volume":"30","author":"R Szeliski","year":"2008","unstructured":"Szeliski, R., Zabih, R., Scharstein, D., Veksler, O., Kolmogorov, V., Agarwala, A., et al. (2008). A comparative study of energy minimization methods for markov random fields with smoothness-based priors. IEEE Transactions on Pattern Analysis and Machine Intelligence, 30(6), 1068\u20131080.","journal-title":"IEEE Transactions on Pattern Analysis and Machine Intelligence"},{"key":"1658_CR60","doi-asserted-by":"publisher","unstructured":"Takeda, Y., Hiura, S., & Sato, K. (2013). Fusing depth from defocus and stereo with coded apertures. In 2013 IEEE conference on computer vision and pattern recognition (pp. 209\u2013216). https:\/\/doi.org\/10.1109\/CVPR.2013.34","DOI":"10.1109\/CVPR.2013.34"},{"key":"1658_CR61","doi-asserted-by":"crossref","unstructured":"Tang, H., Cohen, S., Price, B., Schiller, S., & Kutulakos, K. N. (2017). Depth from defocus in the wild. In IEEE conference on computer vision and pattern recognition (CVPR), pp. 4773\u20134781.","DOI":"10.1109\/CVPR.2017.507"},{"key":"1658_CR62","doi-asserted-by":"crossref","unstructured":"Tao, M. W., Hadap, S., Malik, J., & Ramamoorthi, R. (2013). Depth from combining defocus and correspondence using light-field cameras. In IEEE international conference on computer vision (pp. 673\u2013680).","DOI":"10.1109\/ICCV.2013.89"},{"issue":"3","key":"1658_CR63","doi-asserted-by":"publisher","first-page":"546","DOI":"10.1109\/TPAMI.2016.2554121","volume":"39","author":"MW Tao","year":"2017","unstructured":"Tao, M. W., Srinivasan, P. P., Hadap, S., Malik, J., & Ramamoorthi, R. (2017). Shape estimation from shading, defocus, and correspondence using light-field angular coherence. IEEE Transactions on Pattern Analysis and Machine Intelligence, 39(3), 546\u2013560.","journal-title":"IEEE Transactions on Pattern Analysis and Machine Intelligence"},{"issue":"5","key":"1658_CR64","doi-asserted-by":"publisher","first-page":"815","DOI":"10.1109\/TPAMI.2009.77","volume":"32","author":"E Tola","year":"2010","unstructured":"Tola, E., Lepetit, V., & Fua, P. (2010). Daisy: An efficient dense descriptor applied to wide-baseline stereo. IEEE Transactions on Pattern Analysis and Machine Intelligence, 32(5), 815\u2013830.","journal-title":"IEEE Transactions on Pattern Analysis and Machine Intelligence"},{"issue":"5","key":"1658_CR65","doi-asserted-by":"publisher","first-page":"903","DOI":"10.1007\/s00138-011-0346-8","volume":"23","author":"E Tola","year":"2012","unstructured":"Tola, E., Strecha, C., & Fua, P. (2012). Efficient large-scale multi-view stereo for ultra high-resolution image sets. Machine Vision and Applications, 23(5), 903\u2013920.","journal-title":"Machine Vision and Applications"},{"issue":"12","key":"1658_CR66","doi-asserted-by":"publisher","first-page":"2241","DOI":"10.1109\/TPAMI.2007.70712","volume":"29","author":"G Vogiatzis","year":"2007","unstructured":"Vogiatzis, G., Hernandez, C., Torr, P. H. S., & Cipolla, R. (2007). Multiview stereo via volumetric graph-cuts and occlusion robust photo-consistency. IEEE Transactions on Pattern Analysis and Machine Intelligence, 29(12), 2241\u20132246.","journal-title":"IEEE Transactions on Pattern Analysis and Machine Intelligence"},{"key":"1658_CR67","doi-asserted-by":"crossref","unstructured":"Wang, T.-C., Srikanth, M., & Ramamoorthi, R. (2016). Depth from semi-calibrated stereo and defocus. In IEEE conference on computer vision and pattern recognition (CVPR) (pp. 3717\u20133726).","DOI":"10.1109\/CVPR.2016.404"},{"issue":"3","key":"1658_CR68","doi-asserted-by":"publisher","first-page":"203","DOI":"10.1023\/A:1007905828438","volume":"27","author":"M Watanabe","year":"1998","unstructured":"Watanabe, M., & Nayar, S. (1998). Rational filters for passive depth from defocus. International Journal of Computer Vision, 27(3), 203\u2013225.","journal-title":"International Journal of Computer Vision"},{"key":"1658_CR69","doi-asserted-by":"crossref","unstructured":"Wu, C., Wilburn, B., Matsushita, Y., & Theobalt, C. (2011). High-quality shape from multi-view stereo and shading under general illumination. In CVPR 2011 (pp. 969\u2013976).","DOI":"10.1109\/CVPR.2011.5995388"},{"key":"1658_CR70","doi-asserted-by":"crossref","unstructured":"Yao, Y., Luo, Z., Li, S., Fang, T., & Quan, L. (2018). MVSNet: Depth inference for unstructured multi-view stereo. In Computer Vision\u2013ECCV 2018 (pp. 785\u2013801). Springer, Cham.","DOI":"10.1007\/978-3-030-01237-3_47"},{"key":"1658_CR71","doi-asserted-by":"publisher","unstructured":"Yao, Y., Luo, Z., Li, S., Shen, T., Fang, T., & Quan, L. (2019). Recurrent MVSNet for high-resolution multi-view stereo depth inference. In 2019 IEEE\/CVF conference on computer vision and pattern recognition (CVPR) (pp. 5520\u20135529). https:\/\/doi.org\/10.1109\/CVPR.2019.00567.","DOI":"10.1109\/CVPR.2019.00567"},{"key":"1658_CR72","doi-asserted-by":"crossref","unstructured":"Zagoruyko, S., Komodakis, N. (2015). Learning to compare image patches via convolutional neural networks. In IEEE conference on computer vision and pattern recognition. Proceedings (Vol. 07-12, pp. 43534361\u2013 (2015-06-01)). http:\/\/search.proquest.com\/docview\/1770338817\/.","DOI":"10.1109\/CVPR.2015.7299064"},{"key":"1658_CR73","unstructured":"Zhang, J., Yao, Y., Li, S., Luo, Z., & Fang, T. (2020). Visibility-aware multi-view stereo network. arXiv e-prints arXiv:2008.07928 [cs.CV]."},{"issue":"11","key":"1658_CR74","doi-asserted-by":"publisher","first-page":"1330","DOI":"10.1109\/34.888718","volume":"22","author":"Z Zhang","year":"2000","unstructured":"Zhang, Z. (2000). A flexible new technique for camera calibration. IEEE Transactions on Pattern Analysis and Machine Intelligence, 22(11), 1330\u20131334.","journal-title":"IEEE Transactions on Pattern Analysis and Machine Intelligence"},{"issue":"C","key":"1658_CR75","doi-asserted-by":"publisher","first-page":"47","DOI":"10.1016\/j.isprsjprs.2015.08.008","volume":"109","author":"Z Zhu","year":"2015","unstructured":"Zhu, Z., Stamatopoulos, C., & Fraser, C. S. (2015). Accurate and occlusion-robust multi-view stereo. ISPRS Journal of Photogrammetry and Remote Sensing, 109(C), 47\u201361.","journal-title":"ISPRS Journal of Photogrammetry and Remote Sensing"}],"container-title":["International Journal of Computer Vision"],"original-title":[],"language":"en","link":[{"URL":"https:\/\/link.springer.com\/content\/pdf\/10.1007\/s11263-022-01658-w.pdf","content-type":"application\/pdf","content-version":"vor","intended-application":"text-mining"},{"URL":"https:\/\/link.springer.com\/article\/10.1007\/s11263-022-01658-w\/fulltext.html","content-type":"text\/html","content-version":"vor","intended-application":"text-mining"},{"URL":"https:\/\/link.springer.com\/content\/pdf\/10.1007\/s11263-022-01658-w.pdf","content-type":"application\/pdf","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2022,9,30]],"date-time":"2022-09-30T16:18:23Z","timestamp":1664554703000},"score":1,"resource":{"primary":{"URL":"https:\/\/link.springer.com\/10.1007\/s11263-022-01658-w"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2022,9,13]]},"references-count":75,"journal-issue":{"issue":"11","published-print":{"date-parts":[[2022,11]]}},"alternative-id":["1658"],"URL":"https:\/\/doi.org\/10.1007\/s11263-022-01658-w","relation":{},"ISSN":["0920-5691","1573-1405"],"issn-type":[{"type":"print","value":"0920-5691"},{"type":"electronic","value":"1573-1405"}],"subject":[],"published":{"date-parts":[[2022,9,13]]},"assertion":[{"value":"3 March 2022","order":1,"name":"received","label":"Received","group":{"name":"ArticleHistory","label":"Article History"}},{"value":"16 July 2022","order":2,"name":"accepted","label":"Accepted","group":{"name":"ArticleHistory","label":"Article History"}},{"value":"13 September 2022","order":3,"name":"first_online","label":"First Online","group":{"name":"ArticleHistory","label":"Article History"}}]}}