{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2026,2,10]],"date-time":"2026-02-10T20:09:16Z","timestamp":1770754156776,"version":"3.50.0"},"reference-count":48,"publisher":"Association for Computing Machinery (ACM)","issue":"4","license":[{"start":{"date-parts":[[2020,11,30]],"date-time":"2020-11-30T00:00:00Z","timestamp":1606694400000},"content-version":"vor","delay-in-days":0,"URL":"https:\/\/www.acm.org\/publications\/policies\/copyright_policy#Background"}],"funder":[{"DOI":"10.13039\/501100001809","name":"National Natural Science Fundation of China","doi-asserted-by":"crossref","award":["61971203"],"award-info":[{"award-number":["61971203"]}],"id":[{"id":"10.13039\/501100001809","id-type":"DOI","asserted-by":"crossref"}]},{"name":"National Key R&D Program","award":["2017YFC0806202"],"award-info":[{"award-number":["2017YFC0806202"]}]}],"content-domain":{"domain":["dl.acm.org"],"crossmark-restriction":true},"short-container-title":["ACM Trans. Multimedia Comput. Commun. Appl."],"published-print":{"date-parts":[[2020,11,30]]},"abstract":"<jats:p>Multi-view video plus depth (MVD) is the promising and widely adopted data representation for future 3D visual applications and interactive media. However, compression distortions on depth videos impede the development of such applications, and filters are crucially needed for the quality enhancement at the terminal side. Cross-view priors can intuitively be involved in filter design, but these priors are also distorted in compression and thus the contribution of them can hardly be considered in previous research. In this article, we propose a cross-view optimized filter for depth map quality enhancement by making full use of inner- and cross-view priors. We dedicate to evaluate the contributions of distorted cross-view priors in filtering the current view of depth, and then both inner- and cross-view priors can be involved in the filter design. Thus, distortions of cross-view priors are not barriers again as before. For the purpose of that, mutual information guided cross-view consistency is designed to evaluate the contributions of cross-view priors from compression distortions of MVD. After that, under the framework of global optimization, both inner- and cross-view priors are modeled and taken to minimize the designed energy function where both data accuracy and spatial smoothness are modeled. The experimental results show that the proposed model outperforms state-of-the-art methods, where 3.289 dB and 0.0407 average gains on peak signal-to-noise ratio and structural similarity metrics can be obtained, respectively. For the subjective evaluations, object details and structure information are recovered in the compressed depth video. We also verify our method via several practical applications, including virtual view synthesis for smooth interaction and point cloud for 3D modeling for accuracy evaluation. In these verifications, the ringing and malposition artifacts on object contours are properly handled for interactive video, and discontinuous object surfaces are restored for 3D modeling. All of these results suggest that compression distortions in MVD can be properly filtered by the proposed model, which provides a promising solution for future bandwidth constrained 3D and interactive visual applications.<\/jats:p>","DOI":"10.1145\/3408293","type":"journal-article","created":{"date-parts":[[2020,12,17]],"date-time":"2020-12-17T17:49:26Z","timestamp":1608227366000},"page":"1-19","update-policy":"https:\/\/doi.org\/10.1145\/crossmark-policy","source":"Crossref","is-referenced-by-count":4,"title":["Make Full Use of Priors"],"prefix":"10.1145","volume":"16","author":[{"given":"Xin","family":"He","sequence":"first","affiliation":[{"name":"Huazhong University of Science and Technology, Wuhan, China"}]},{"given":"Qiong","family":"Liu","sequence":"additional","affiliation":[{"name":"Huazhong University of Science and Technology, Wuhan, China"}]},{"ORCID":"https:\/\/orcid.org\/0000-0002-5695-1046","authenticated-orcid":false,"given":"You","family":"Yang","sequence":"additional","affiliation":[{"name":"Huazhong University of Science and Technology, Wuhan, China"}]}],"member":"320","published-online":{"date-parts":[[2020,12,17]]},"reference":[{"key":"e_1_2_1_1_1","volume-title":"Proceedings of the IEEE International Conference on Computer Vision. 3828--3838","author":"Godard Cl\u00e9ment"},{"key":"e_1_2_1_2_1","doi-asserted-by":"publisher","DOI":"10.1109\/TPAMI.2019.2894422"},{"key":"e_1_2_1_3_1","volume-title":"Proceedings of the 7th Meeting of the JCT.","author":"M\u00fcller Karsten","year":"2014"},{"key":"e_1_2_1_4_1","unstructured":"Guillaume Rochette Chris Russell and Richard Bowden. 2019. Weakly-supervised 3D pose estimation from a single image using multi-view consistency. arXiv:1909.06119  Guillaume Rochette Chris Russell and Richard Bowden. 2019. Weakly-supervised 3D pose estimation from a single image using multi-view consistency. arXiv:1909.06119"},{"key":"e_1_2_1_5_1","doi-asserted-by":"publisher","DOI":"10.1109\/TMM.2011.2169045"},{"key":"e_1_2_1_6_1","doi-asserted-by":"publisher","DOI":"10.1109\/TPAMI.2012.120"},{"key":"e_1_2_1_7_1","doi-asserted-by":"publisher","DOI":"10.1109\/ICIP.2010.5650661"},{"key":"e_1_2_1_8_1","doi-asserted-by":"publisher","DOI":"10.1109\/TMM.2012.2229264"},{"key":"e_1_2_1_9_1","doi-asserted-by":"publisher","DOI":"10.1109\/34.969114"},{"key":"e_1_2_1_10_1","volume-title":"Proceedings of the Workshop on Multi-Camera and Multi-Modal Sensor Fusion Algorithms and Applications.","author":"Chan Derek","year":"2008"},{"key":"e_1_2_1_11_1","doi-asserted-by":"publisher","DOI":"10.1109\/DCC.2019.00072"},{"key":"e_1_2_1_12_1","doi-asserted-by":"publisher","DOI":"10.1109\/TCSVT.2013.2278160"},{"key":"e_1_2_1_13_1","volume-title":"Proceedings of the IEEE International Conference on Communications. 143--147","author":"Dai Rui"},{"key":"e_1_2_1_14_1","unstructured":"James Diebel and Sebastian Thrun. 2006. An application of Markov random fields to range sensing. In Advances in Neural Information Processing Systems. 291--298.  James Diebel and Sebastian Thrun. 2006. An application of Markov random fields to range sensing. In Advances in Neural Information Processing Systems. 291--298."},{"key":"e_1_2_1_15_1","doi-asserted-by":"publisher","DOI":"10.1109\/TMM.2016.2613824"},{"key":"e_1_2_1_16_1","unstructured":"David Eigen Christian Puhrsch and Rob Fergus. 2014. Depth map prediction from a single image using a multi-scale deep network. In Advances in Neural Information Processing Systems. 2366--2374.  David Eigen Christian Puhrsch and Rob Fergus. 2014. Depth map prediction from a single image using a multi-scale deep network. In Advances in Neural Information Processing Systems. 2366--2374."},{"key":"e_1_2_1_17_1","doi-asserted-by":"publisher","DOI":"10.1109\/JSTSP.2010.2052783"},{"key":"e_1_2_1_18_1","doi-asserted-by":"publisher","DOI":"10.1109\/3DTV.2007.4379449"},{"key":"e_1_2_1_19_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR.2015.7299115"},{"key":"e_1_2_1_20_1","doi-asserted-by":"publisher","DOI":"10.1109\/TPAMI.2012.213"},{"key":"e_1_2_1_21_1","doi-asserted-by":"publisher","DOI":"10.1109\/ICPR.2010.579"},{"key":"e_1_2_1_22_1","doi-asserted-by":"publisher","DOI":"10.1109\/ICMEW.2015.7169850"},{"key":"e_1_2_1_23_1","volume-title":"Three-Dimensional Image Processing (3DIP) and Applications","author":"Kim Deukhyeon"},{"key":"e_1_2_1_24_1","doi-asserted-by":"publisher","DOI":"10.1145\/1276377.1276497"},{"key":"e_1_2_1_25_1","doi-asserted-by":"publisher","DOI":"10.1016\/j.neucom.2012.09.009"},{"key":"e_1_2_1_26_1","doi-asserted-by":"publisher","DOI":"10.1109\/LSP.2012.2190060"},{"key":"e_1_2_1_27_1","doi-asserted-by":"publisher","DOI":"10.1109\/TIP.2016.2612826"},{"key":"e_1_2_1_28_1","doi-asserted-by":"publisher","DOI":"10.1109\/VCIP.2016.7805550"},{"key":"e_1_2_1_29_1","volume-title":"Proceedings of the IEEE International Conference on Acoustics, Speech, and Signal Processing (ICASSP\u201911)","author":"Lu Jiangbo"},{"key":"e_1_2_1_30_1","volume-title":"Do","author":"Min Dongbo","year":"2012"},{"key":"e_1_2_1_31_1","doi-asserted-by":"publisher","DOI":"10.1109\/TMM.2011.2128862"},{"key":"e_1_2_1_32_1","doi-asserted-by":"publisher","DOI":"10.1109\/ICCV.2011.6126423"},{"key":"e_1_2_1_33_1","doi-asserted-by":"publisher","DOI":"10.1109\/TMI.2003.815867"},{"key":"e_1_2_1_34_1","doi-asserted-by":"publisher","DOI":"10.1109\/TMM.2018.2845699"},{"key":"e_1_2_1_35_1","doi-asserted-by":"publisher","DOI":"10.1109\/TMM.2013.2246148"},{"key":"e_1_2_1_36_1","volume-title":"San Jose","author":"M\u00fcller Karsten","year":"2014"},{"key":"e_1_2_1_37_1","doi-asserted-by":"publisher","DOI":"10.1109\/JSTSP.2013.2283657"},{"key":"e_1_2_1_38_1","volume-title":"Busan","author":"Tanimoto M.","year":"2008"},{"key":"e_1_2_1_39_1","volume-title":"Proceedings of the IEEE International Conference on Computer Vision (ICCV\u201998)","author":"Tomasi C."},{"key":"e_1_2_1_40_1","doi-asserted-by":"publisher","DOI":"10.1109\/ACCESS.2020.2981161"},{"key":"e_1_2_1_41_1","doi-asserted-by":"publisher","DOI":"10.1007\/s00371-013-0896-z"},{"key":"e_1_2_1_42_1","doi-asserted-by":"publisher","DOI":"10.1109\/TIP.2003.819861"},{"key":"e_1_2_1_43_1","doi-asserted-by":"publisher","DOI":"10.1109\/TMM.2015.2457678"},{"key":"e_1_2_1_44_1","doi-asserted-by":"publisher","DOI":"10.1016\/j.image.2013.12.005"},{"key":"e_1_2_1_45_1","doi-asserted-by":"crossref","unstructured":"J. Yang X. Ye K. Li C. Hou and Y. Wang. 2014. Color-guided depth recovery from RGB-D data using an adaptive autoregressive model.IEEE Transactions on Image Processing 23 8 (2014) 3443--3458.  J. Yang X. Ye K. Li C. Hou and Y. Wang. 2014. Color-guided depth recovery from RGB-D data using an adaptive autoregressive model.IEEE Transactions on Image Processing 23 8 (2014) 3443--3458.","DOI":"10.1109\/TIP.2014.2329776"},{"key":"e_1_2_1_46_1","doi-asserted-by":"publisher","DOI":"10.1109\/TIP.2018.2867740"},{"key":"e_1_2_1_47_1","doi-asserted-by":"publisher","DOI":"10.1016\/j.image.2017.02.009"},{"key":"e_1_2_1_48_1","doi-asserted-by":"publisher","DOI":"10.1049\/el.2014.3912"}],"container-title":["ACM Transactions on Multimedia Computing, Communications, and Applications"],"original-title":[],"language":"en","link":[{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/3408293","content-type":"unspecified","content-version":"vor","intended-application":"text-mining"},{"URL":"https:\/\/dl.acm.org\/doi\/pdf\/10.1145\/3408293","content-type":"unspecified","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2025,6,17]],"date-time":"2025-06-17T22:39:01Z","timestamp":1750199941000},"score":1,"resource":{"primary":{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/3408293"}},"subtitle":["Cross-View Optimized Filter for Multi-View Depth Enhancement"],"short-title":[],"issued":{"date-parts":[[2020,11,30]]},"references-count":48,"journal-issue":{"issue":"4","published-print":{"date-parts":[[2020,11,30]]}},"alternative-id":["10.1145\/3408293"],"URL":"https:\/\/doi.org\/10.1145\/3408293","relation":{},"ISSN":["1551-6857","1551-6865"],"issn-type":[{"value":"1551-6857","type":"print"},{"value":"1551-6865","type":"electronic"}],"subject":[],"published":{"date-parts":[[2020,11,30]]},"assertion":[{"value":"2019-10-01","order":0,"name":"received","label":"Received","group":{"name":"publication_history","label":"Publication History"}},{"value":"2020-06-01","order":1,"name":"accepted","label":"Accepted","group":{"name":"publication_history","label":"Publication History"}},{"value":"2020-12-17","order":2,"name":"published","label":"Published","group":{"name":"publication_history","label":"Publication History"}}]}}