{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2026,2,24]],"date-time":"2026-02-24T16:25:49Z","timestamp":1771950349424,"version":"3.50.1"},"reference-count":42,"publisher":"MDPI AG","issue":"9","license":[{"start":{"date-parts":[[2021,5,9]],"date-time":"2021-05-09T00:00:00Z","timestamp":1620518400000},"content-version":"vor","delay-in-days":0,"URL":"https:\/\/creativecommons.org\/licenses\/by\/4.0\/"}],"funder":[{"name":"technology project of 630 Hainan province of China","award":["ZDKJ202017"],"award-info":[{"award-number":["ZDKJ202017"]}]},{"DOI":"10.13039\/100007834","name":"Natural Science Foundation of Ningbo","doi-asserted-by":"publisher","award":["61701463"],"award-info":[{"award-number":["61701463"]}],"id":[{"id":"10.13039\/100007834","id-type":"DOI","asserted-by":"publisher"}]}],"content-domain":{"domain":[],"crossmark-restriction":false},"short-container-title":["Sensors"],"abstract":"<jats:p>As one of the key requirements for underwater exploration, underwater depth map estimation is of great importance in underwater vision research. Although significant progress has been achieved in the fields of image-to-image translation and depth map estimation, a gap between normal depth map estimation and underwater depth map estimation still remains. Additionally, it is a great challenge to build a mapping function that converts a single underwater image into an underwater depth map due to the lack of paired data. Moreover, the ever-changing underwater environment further intensifies the difficulty of finding an optimal mapping solution. To eliminate these bottlenecks, we developed a novel image-to-image framework for underwater image synthesis and depth map estimation in underwater conditions. For the problem of the lack of paired data, by translating hazy in-air images (with a depth map) into underwater images, we initially obtained a paired dataset of underwater images and corresponding depth maps. To enrich our synthesized underwater dataset, we further translated hazy in-air images into a series of continuously changing underwater images with a specified style. For the depth map estimation, we included a coarse-to-fine network to provide a precise depth map estimation result. We evaluated the efficiency of our framework for a real underwater RGB-D dataset. The experimental results show that our method can provide a diversity of underwater images and the best depth map estimation precision.<\/jats:p>","DOI":"10.3390\/s21093268","type":"journal-article","created":{"date-parts":[[2021,5,10]],"date-time":"2021-05-10T02:54:58Z","timestamp":1620615298000},"page":"3268","update-policy":"https:\/\/doi.org\/10.3390\/mdpi_crossmark_policy","source":"Crossref","is-referenced-by-count":10,"title":["Unpaired Underwater Image Synthesis with a Disentangled Representation for Underwater Depth Map Prediction"],"prefix":"10.3390","volume":"21","author":[{"given":"Qi","family":"Zhao","sequence":"first","affiliation":[{"name":"College of Information Science and Engineering, Ocean University of China, Qingdao 266100, China"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Zhichao","family":"Xin","sequence":"additional","affiliation":[{"name":"College of Information Science and Engineering, Ocean University of China, Qingdao 266100, China"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"ORCID":"https:\/\/orcid.org\/0000-0003-4372-1767","authenticated-orcid":false,"given":"Zhibin","family":"Yu","sequence":"additional","affiliation":[{"name":"College of Information Science and Engineering, Ocean University of China, Qingdao 266100, China"},{"name":"Sanya Oceanographic Institution, Ocean University of China, Sanya 572000, China"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Bing","family":"Zheng","sequence":"additional","affiliation":[{"name":"College of Information Science and Engineering, Ocean University of China, Qingdao 266100, China"},{"name":"Sanya Oceanographic Institution, Ocean University of China, Sanya 572000, China"}],"role":[{"role":"author","vocabulary":"crossref"}]}],"member":"1968","published-online":{"date-parts":[[2021,5,9]]},"reference":[{"key":"ref_1","first-page":"2366","article-title":"Depth map prediction from a single image using a multi-scale deep network","volume":"27","author":"Eigen","year":"2014","journal-title":"Adv. Neural Inf. Process. Syst."},{"key":"ref_2","doi-asserted-by":"crossref","unstructured":"Chi, S., Xie, Z., and Chen, W. (2016). A laser line auto-scanning system for underwater 3D reconstruction. Sensors, 16.","DOI":"10.3390\/s16091534"},{"key":"ref_3","doi-asserted-by":"crossref","unstructured":"Palomer, A., Ridao, P., Ribas, D., and Forest, J. (2017). Underwater 3D laser scanners: The deformation of the plane. Sensing and Control for Autonomous Vehicles, Springer.","DOI":"10.1007\/978-3-319-55372-6_4"},{"key":"ref_4","doi-asserted-by":"crossref","first-page":"75","DOI":"10.4031\/MTSJ.51.1.8","article-title":"Review of underwater machine vision technology and its applications","volume":"51","author":"Xi","year":"2017","journal-title":"Mar. Technol. Soc. J."},{"key":"ref_5","doi-asserted-by":"crossref","unstructured":"Dancu, A., Fourgeaud, M., Franjcic, Z., and Avetisyan, R. (2014). Underwater reconstruction using depth sensors. SIGGRAPH ASIA Technical Briefs, ACM.","DOI":"10.1145\/2669024.2669042"},{"key":"ref_6","doi-asserted-by":"crossref","unstructured":"Churnside, J.H., Marchbanks, R.D., Lembke, C., and Beckler, J. (2017). Optical backscattering measured by airborne lidar and underwater glider. Remote Sens., 9.","DOI":"10.3390\/rs9040379"},{"key":"ref_7","doi-asserted-by":"crossref","first-page":"231","DOI":"10.5194\/isprs-archives-XLII-2-W3-231-2017","article-title":"Depth cameras on UAVs: A first approach","volume":"42","author":"Deris","year":"2017","journal-title":"Int. Arch. Photogramm. Remote Sens. Spat. Inf. Sci."},{"key":"ref_8","doi-asserted-by":"crossref","first-page":"1587","DOI":"10.1049\/iet-ipr.2019.0117","article-title":"Review of underwater image restoration algorithms","volume":"13","author":"Ahamed","year":"2019","journal-title":"IET Image Process."},{"key":"ref_9","doi-asserted-by":"crossref","first-page":"31525","DOI":"10.3390\/s151229864","article-title":"Optical sensors and methods for underwater 3D reconstruction","volume":"15","year":"2015","journal-title":"Sensors"},{"key":"ref_10","doi-asserted-by":"crossref","first-page":"54241","DOI":"10.1109\/ACCESS.2018.2870854","article-title":"The synthesis of unpaired underwater images using a multistyle generative adversarial network","volume":"6","author":"Li","year":"2018","journal-title":"IEEE Access"},{"key":"ref_11","doi-asserted-by":"crossref","unstructured":"Gupta, H., and Mitra, K. (2019, January 22\u201325). Unsupervised single image underwater depth estimation. Proceedings of the IEEE International Conference on Image Processing (ICIP), Taipei, Taiwan.","DOI":"10.1109\/ICIP.2019.8804200"},{"key":"ref_12","doi-asserted-by":"crossref","unstructured":"Zhu, J.Y., Park, T., Isola, P., and Efros, A.A. (2017, January 22\u201329). Unpaired image-to-image translation using cycle-consistent adversarial networks. Proceedings of the IEEE International Conference on Computer Vision, Venice, Italy.","DOI":"10.1109\/ICCV.2017.244"},{"key":"ref_13","doi-asserted-by":"crossref","unstructured":"Isola, P., Zhu, J.Y., Zhou, T., and Efros, A.A. (2017, January 21\u201326). Image-to-image translation with conditional adversarial networks. Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition, Honolulu, HI, USA.","DOI":"10.1109\/CVPR.2017.632"},{"key":"ref_14","doi-asserted-by":"crossref","unstructured":"Wang, T.C., Liu, M.Y., Zhu, J.Y., Tao, A., Kautz, J., and Catanzaro, B. (2018, January 18\u201322). High-resolution image synthesis and semantic manipulation with conditional GANs. Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition, Salt Lake City, UT, USA.","DOI":"10.1109\/CVPR.2018.00917"},{"key":"ref_15","doi-asserted-by":"crossref","unstructured":"Choi, Y., Choi, M., Kim, M., Ha, J.W., Kim, S., and Choo, J. (2018, January 18\u201322). StarGAN: Unified generative adversarial networks for multi-domain image-to-image translation. Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition, Salt Lake City, UT, USA.","DOI":"10.1109\/CVPR.2018.00916"},{"key":"ref_16","first-page":"387","article-title":"WaterGAN: Unsupervised generative network to enable real-time color correction of monocular underwater images","volume":"3","author":"Li","year":"2017","journal-title":"IEEE Robot. Autom. Lett."},{"key":"ref_17","doi-asserted-by":"crossref","unstructured":"Cao, K., Peng, Y.T., and Cosman, P.C. (2018, January 8\u201310). Underwater image restoration using deep networks to estimate background light and scene depth. Proceedings of the IEEE Southwest Symposium on Image Analysis and Interpretation, Las Vegas, NV, USA.","DOI":"10.1109\/SSIAI.2018.8470347"},{"key":"ref_18","doi-asserted-by":"crossref","first-page":"107038","DOI":"10.1016\/j.patcog.2019.107038","article-title":"Underwater scene prior inspired deep underwater image and video enhancement","volume":"98","author":"Li","year":"2020","journal-title":"Pattern Recognit."},{"key":"ref_19","doi-asserted-by":"crossref","unstructured":"Skinner, K.A., Zhang, J., Olson, E.A., and Johnson-Roberson, M. (2019, January 20\u201324). Uwstereonet: Unsupervised learning for depth estimation and color correction of underwater stereo imagery. Proceedings of the International Conference on Robotics and Automation, Montreal, QC, Canada.","DOI":"10.1109\/ICRA.2019.8794272"},{"key":"ref_20","unstructured":"Wang, N., Zhou, Y., Han, F., Zhu, H., and Zheng, Y. (2019). UWGAN: Underwater GAN for real-world underwater color restoration and dehazing. arXiv."},{"key":"ref_21","unstructured":"Chen, X., Duan, Y., Houthooft, R., Schulman, J., Sutskever, I., and Abbeel, P. (2016). InfoGAN: Interpretable representation learning by information maximizing generative adversarial nets. arXiv."},{"key":"ref_22","doi-asserted-by":"crossref","unstructured":"Spurr, A., Aksan, E., and Hilliges, O. (2017). Guiding InfoGAN with semi-supervision. Joint European Conference on Machine Learning and Knowledge Discovery in Databases, Springer.","DOI":"10.1007\/978-3-319-71249-9_8"},{"key":"ref_23","doi-asserted-by":"crossref","unstructured":"Eigen, D., and Fergus, R. (2015, January 13\u201316). Predicting depth, surface normals and semantic labels with a common multi-scale convolutional architecture. Proceedings of the IEEE International Conference on Computer Vision, Santiago, Chile.","DOI":"10.1109\/ICCV.2015.304"},{"key":"ref_24","doi-asserted-by":"crossref","unstructured":"Huang, G., Liu, Z., Van Der Maaten, L., and Weinberger, K.Q. (2017, January 21\u201326). Densely connected convolutional networks. Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition, Honolulu, HI, USA.","DOI":"10.1109\/CVPR.2017.243"},{"key":"ref_25","unstructured":"Goodfellow, I.J., Pouget-Abadie, J., Mirza, M., Xu, B., Warde-Farley, D., Ozair, S., Courville, A., and Bengio, Y. (2014, January 8\u201313). Generative adversarial nets. Proceedings of the International Conference on Neural Information Processing Systems, Montreal, QC, Canada."},{"key":"ref_26","doi-asserted-by":"crossref","unstructured":"Mao, X., Li, Q., Xie, H., Lau, R.Y., Wang, Z., and Paul Smolley, S. (2017, January 22\u201329). Least squares generative adversarial networks. Proceedings of the IEEE International Conference on Computer Vision, Venice, Italy.","DOI":"10.1109\/ICCV.2017.304"},{"key":"ref_27","doi-asserted-by":"crossref","first-page":"4066","DOI":"10.1109\/TIP.2018.2836316","article-title":"Perceptual adversarial networks for image-to-image transformation","volume":"27","author":"Wang","year":"2018","journal-title":"IEEE Trans. Image Process."},{"key":"ref_28","unstructured":"Simonyan, K., and Zisserman, A. (2015, January 7\u20139). Very deep convolutional networks for large-scale image recognition. Proceedings of the International Conference on Learning Representations, San Diego, CA, USA."},{"key":"ref_29","doi-asserted-by":"crossref","unstructured":"Kupyn, O., Martyniuk, T., Wu, J., and Wang, Z. (2019, January 27\u201329). DeblurGAN-v2: Deblurring (orders-of-magnitude) faster and better. Proceedings of the IEEE\/CVF International Conference on Computer Vision, Seoul, Korea.","DOI":"10.1109\/ICCV.2019.00897"},{"key":"ref_30","doi-asserted-by":"crossref","unstructured":"Choi, Y., Uh, Y., Yoo, J., and Ha, J.W. (2020, January 13\u201319). StarGAN v2: Diverse image synthesis for multiple domains. Proceedings of the IEEE\/CVF Conference on Computer Vision and Pattern Recognition, Seattle, WA, USA.","DOI":"10.1109\/CVPR42600.2020.00821"},{"key":"ref_31","doi-asserted-by":"crossref","unstructured":"Hu, J., Ozay, M., Zhang, Y., and Okatani, T. (2019, January 7\u201311). Revisiting single image depth estimation: Toward higher resolution maps with accurate object boundaries. Proceedings of the IEEE Winter Conference on Applications of Computer Vision (WACV), Waikoloa Village, HI, USA.","DOI":"10.1109\/WACV.2019.00116"},{"key":"ref_32","first-page":"122","article-title":"GAN generation of synthetic multispectral satellite images","volume":"Volume 11533","author":"Abady","year":"2020","journal-title":"Image and Signal Processing for Remote Sensing XXVI"},{"key":"ref_33","first-page":"2341","article-title":"Single image haze removal using dark channel prior","volume":"33","author":"He","year":"2010","journal-title":"IEEE Trans. Pattern Anal. Mach. Intell."},{"key":"ref_34","doi-asserted-by":"crossref","first-page":"24","DOI":"10.1109\/MCG.2016.26","article-title":"Underwater depth estimation and image restoration based on single images","volume":"36","author":"Drews","year":"2016","journal-title":"IEEE Comput. Graph. Appl."},{"key":"ref_35","unstructured":"Berman, D., Treibitz, T., and Avidan, S. (2017, January 9\u201312). Diving into haze-lines: Color restoration of underwater images. Proceedings of the British Machine Vision Conference (BMVC), London, UK."},{"key":"ref_36","doi-asserted-by":"crossref","unstructured":"Ancuti, C., Ancuti, C.O., and De Vleeschouwer, C. (2016, January 25\u201328). D-hazy: A dataset to evaluate quantitatively dehazing algorithms. Proceedings of the IEEE International Conference on Image Processing (ICIP), Phoenix, AZ, USA.","DOI":"10.1109\/ICIP.2016.7532754"},{"key":"ref_37","doi-asserted-by":"crossref","unstructured":"Xiao, J., Hays, J., Ehinger, K.A., Oliva, A., and Torralba, A. (2010, January 13\u201318). Sun database: Large-scale scene recognition from abbey to zoo. Proceedings of the IEEE Computer Society Conference on Computer Vision and Pattern Recognition, San Francisco, CA, USA.","DOI":"10.1109\/CVPR.2010.5539970"},{"key":"ref_38","unstructured":"Miyato, T., Kataoka, T., Koyama, M., and Yoshida, Y. (30\u20133, January 30). Spectral normalization for generative adversarial networks. Proceedings of the International Conference on Learning Representations, Vancouver, BC, Canada."},{"key":"ref_39","unstructured":"Brock, A., Donahue, J., and Simonyan, K. (30\u20133, January 30). Large scale GAN training for high fidelity natural image synthesis. Proceedings of the International Conference on Learning Representations, Vancouver, BC, Canada."},{"key":"ref_40","unstructured":"Zhang, H., Goodfellow, I., Metaxas, D., and Odena, A. (2019, January 9\u201315). Self-attention generative adversarial networks. Proceedings of the International Conference on Machine Learning, Long Beach, CA, USA."},{"key":"ref_41","doi-asserted-by":"crossref","unstructured":"Silberman, N., Hoiem, D., Kohli, P., and Fergus, R. (2012). Indoor segmentation and support inference from rgbd images. European Conference on Computer Vision, Springer.","DOI":"10.1007\/978-3-642-33715-4_54"},{"key":"ref_42","doi-asserted-by":"crossref","unstructured":"Chen, Y., Li, W., Sakaridis, C., Dai, D., and Van Gool, L. (2018, January 18\u201322). Domain adaptive faster R-CNN for object detection in the wild. Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition, Salt Lake City, UT, USA.","DOI":"10.1109\/CVPR.2018.00352"}],"container-title":["Sensors"],"original-title":[],"language":"en","link":[{"URL":"https:\/\/www.mdpi.com\/1424-8220\/21\/9\/3268\/pdf","content-type":"unspecified","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2025,10,11]],"date-time":"2025-10-11T05:58:28Z","timestamp":1760162308000},"score":1,"resource":{"primary":{"URL":"https:\/\/www.mdpi.com\/1424-8220\/21\/9\/3268"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2021,5,9]]},"references-count":42,"journal-issue":{"issue":"9","published-online":{"date-parts":[[2021,5]]}},"alternative-id":["s21093268"],"URL":"https:\/\/doi.org\/10.3390\/s21093268","relation":{},"ISSN":["1424-8220"],"issn-type":[{"value":"1424-8220","type":"electronic"}],"subject":[],"published":{"date-parts":[[2021,5,9]]}}}