{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2026,7,23]],"date-time":"2026-07-23T19:47:25Z","timestamp":1784836045192,"version":"3.55.0"},"reference-count":75,"publisher":"MDPI AG","issue":"9","license":[{"start":{"date-parts":[[2022,5,9]],"date-time":"2022-05-09T00:00:00Z","timestamp":1652054400000},"content-version":"vor","delay-in-days":0,"URL":"https:\/\/creativecommons.org\/licenses\/by\/4.0\/"}],"funder":[{"DOI":"10.13039\/501100003052","name":"Ministry of Trade, Industry &amp; Energy (MOTIE, Korea)","doi-asserted-by":"publisher","award":["20013726"],"award-info":[{"award-number":["20013726"]}],"id":[{"id":"10.13039\/501100003052","id-type":"DOI","asserted-by":"publisher"}]}],"content-domain":{"domain":[],"crossmark-restriction":false},"short-container-title":["Sensors"],"abstract":"<jats:p>The salient object detection (SOD) technology predicts which object will attract the attention of an observer surveying a particular scene. Most state-of-the-art SOD methods are top-down mechanisms that apply fully convolutional networks (FCNs) of various structures to RGB images, extract features from them, and train a network. However, owing to the variety of factors that affect visual saliency, securing sufficient features from a single color space is difficult. Therefore, in this paper, we propose a multi-color space network (MCSNet) to detect salient objects using various saliency cues. First, the images were converted to HSV and grayscale color spaces to obtain saliency cues other than those provided by RGB color information. Each saliency cue was fed into two parallel VGG backbone networks to extract features. Contextual information was obtained from the extracted features using atrous spatial pyramid pooling (ASPP). The features obtained from both paths were passed through the attention module, and channel and spatial features were highlighted. Finally, the final saliency map was generated using a step-by-step residual refinement module (RRM). Furthermore, the network was trained with a bidirectional loss to supervise saliency detection results. Experiments on five public benchmark datasets showed that our proposed network achieved superior performance in terms of both subjective results and objective metrics.<\/jats:p>","DOI":"10.3390\/s22093588","type":"journal-article","created":{"date-parts":[[2022,5,10]],"date-time":"2022-05-10T00:30:28Z","timestamp":1652142628000},"page":"3588","update-policy":"https:\/\/doi.org\/10.3390\/mdpi_crossmark_policy","source":"Crossref","is-referenced-by-count":5,"title":["Multi-Color Space Network for Salient Object Detection"],"prefix":"10.3390","volume":"22","author":[{"ORCID":"https:\/\/orcid.org\/0000-0001-5728-9368","authenticated-orcid":false,"given":"Kyungjun","family":"Lee","sequence":"first","affiliation":[{"name":"Department of Electronics and Computer Engineering, Hanyang University, Seoul 04763, Korea"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"ORCID":"https:\/\/orcid.org\/0000-0002-3759-3116","authenticated-orcid":false,"given":"Jechang","family":"Jeong","sequence":"additional","affiliation":[{"name":"Department of Electronics and Computer Engineering, Hanyang University, Seoul 04763, Korea"}],"role":[{"vocabulary":"crossref","role":"author"}]}],"member":"1968","published-online":{"date-parts":[[2022,5,9]]},"reference":[{"key":"ref_1","doi-asserted-by":"crossref","unstructured":"Li, J., and Gao, W. (2014). Visual Saliency Computation: A Machine Learning Perspective, Springer.","DOI":"10.1007\/978-3-319-05642-5"},{"key":"ref_2","doi-asserted-by":"crossref","unstructured":"Donoser, M., Urschler, M., Hirzer, M., and Bischof, H. (October, January 29). Saliency driven total variation segmentation. Proceedings of the 2009 IEEE 12th International Conference on Computer Vision, Kyoto, Japan.","DOI":"10.1109\/ICCV.2009.5459296"},{"key":"ref_3","doi-asserted-by":"crossref","first-page":"4290","DOI":"10.1109\/TIP.2012.2199502","article-title":"3-D object retrieval and recognition with hypergraph analysis","volume":"21","author":"Gao","year":"2012","journal-title":"IEEE Trans. Image Process."},{"key":"ref_4","doi-asserted-by":"crossref","unstructured":"Borji, A., Frintrop, S., Sihite, D.N., and Itti, L. (2012, January 16\u201321). Adaptive object tracking by learning background context. Proceedings of the 2012 IEEE Computer Society Conference on Computer Vision and Pattern Recognition Workshops, Providence, RI, USA.","DOI":"10.1109\/CVPRW.2012.6239191"},{"key":"ref_5","doi-asserted-by":"crossref","first-page":"300","DOI":"10.1109\/TPAMI.2007.40","article-title":"Rapid biologically-inspired scene classification using features shared with visual attention","volume":"29","author":"Siagian","year":"2007","journal-title":"IEEE Trans. Pattern Anal. Mach. Intell."},{"key":"ref_6","first-page":"185","article-title":"A novel multiresolution spatiotemporal saliency detection model and its applications in image and video compression","volume":"19","author":"Guo","year":"2009","journal-title":"IEEE Trans. Image Process."},{"key":"ref_7","unstructured":"Lee, K.J., Wee, S.W., and Jeong, J.C. (2017, January 19\u201320). Pre-filtering with Contents-based Adaptive Filter Set for High Efficiency Video Coding Standard. Proceedings of the IEIE International Conference on Electronics, Information, and Communication 2017, Piscataway, NJ, USA."},{"key":"ref_8","doi-asserted-by":"crossref","first-page":"509","DOI":"10.1177\/1073858413514136","article-title":"Bottom-up and top-down attention: Different processes and overlapping neural systems","volume":"20","author":"Katsuki","year":"2014","journal-title":"Neuroscientist"},{"key":"ref_9","doi-asserted-by":"crossref","unstructured":"Wolfe, J.M. (2014). Guidance of visual search by preattentive information. Neurobiol. Atten., 101\u2013104.","DOI":"10.1016\/B978-012375731-9\/50021-5"},{"key":"ref_10","doi-asserted-by":"crossref","first-page":"1254","DOI":"10.1109\/34.730558","article-title":"A model of saliency-based visual attention for rapid scene analysis","volume":"20","author":"Itti","year":"1998","journal-title":"IEEE Trans. Pattern Anal. Mach. Intell."},{"key":"ref_11","doi-asserted-by":"crossref","unstructured":"Achanta, R., Estrada, F., Wils, P., and S\u00fcsstrunk, S. (2008). Salient region detection and segmentation. Proceedings of the International Conference on Computer Vision Systems, Springer.","DOI":"10.1007\/978-3-540-79547-6_7"},{"key":"ref_12","doi-asserted-by":"crossref","first-page":"1915","DOI":"10.1109\/TPAMI.2011.272","article-title":"Context-aware saliency detection","volume":"34","author":"Goferman","year":"2011","journal-title":"IEEE Trans. Pattern Anal. Mach. Intell."},{"key":"ref_13","doi-asserted-by":"crossref","first-page":"569","DOI":"10.1109\/TPAMI.2014.2345401","article-title":"Global contrast based salient region detection","volume":"37","author":"Cheng","year":"2015","journal-title":"IEEE Trans. Pattern Anal. Mach. Intell."},{"key":"ref_14","doi-asserted-by":"crossref","unstructured":"Liu, Z., Meur, L., and Luo, S. (2013, January 3\u20135). Superpixel-based saliency detection. Proceedings of the 2013 14th International Workshop on Image Analysis for Multimedia Interactive Services (WIAMIS), Paris, France.","DOI":"10.1109\/WIAMIS.2013.6616119"},{"key":"ref_15","doi-asserted-by":"crossref","first-page":"194","DOI":"10.1038\/35058500","article-title":"Computational modelling of visual attention","volume":"2","author":"Itti","year":"2001","journal-title":"Nat. Rev. Neurosci."},{"key":"ref_16","doi-asserted-by":"crossref","first-page":"210","DOI":"10.1016\/j.tins.2011.02.003","article-title":"Mechanisms of top-down attention","volume":"34","author":"Baluch","year":"2011","journal-title":"Trends Neurosci."},{"key":"ref_17","doi-asserted-by":"crossref","first-page":"2278","DOI":"10.1109\/5.726791","article-title":"Gradient-based learning applied to document recognition","volume":"86","author":"LeCun","year":"1998","journal-title":"Proc. IEEE"},{"key":"ref_18","doi-asserted-by":"crossref","unstructured":"Long, J., Shelhamer, E., and Darrell, T. (2015, January 7\u201312). Fully convolutional networks for semantic segmentation. Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition, Boston, MA, USA.","DOI":"10.1109\/CVPR.2015.7298965"},{"key":"ref_19","doi-asserted-by":"crossref","unstructured":"Wang, L., Wang, L., Lu, H., Zhang, P., and Ruan, X. (2016, January 11\u201314). Saliency detection with recurrent fully convolutional networks. Proceedings of the European Conference on Computer Vision, Amsterdam, The Netherlands.","DOI":"10.1007\/978-3-319-46493-0_50"},{"key":"ref_20","doi-asserted-by":"crossref","unstructured":"Zhang, P., Wang, D., Lu, H., Wang, H., and Ruan, X. (2017, January 22\u201329). Amulet: Aggregating multi-level convolutional features for salient object detection. Proceedings of the IEEE International Conference on Computer Vision, Venice, Italy.","DOI":"10.1109\/ICCV.2017.31"},{"key":"ref_21","doi-asserted-by":"crossref","unstructured":"Liu, N., Han, J., and Yang, M.H. (2018, January 18\u201323). Picanet: Learning pixel-wise contextual attention for saliency detection. Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition, Salt Lake City, UT, USA.","DOI":"10.1109\/CVPR.2018.00326"},{"key":"ref_22","doi-asserted-by":"crossref","unstructured":"Liu, J.J., Hou, Q., Cheng, M.M., Feng, J., and Jiang, J. (2019, January 15\u201320). A simple pooling-based design for real-time salient object detection. Proceedings of the IEEE\/CVF Conference on Computer Vision and Pattern Recognition, Long Beach, CA, USA.","DOI":"10.1109\/CVPR.2019.00404"},{"key":"ref_23","unstructured":"Wei, J., Wang, S., and Huang, Q. (2020, January 7\u201312). F3Net: Fusion, feedback and focus for salient object detection. Proceedings of the AAAI Conference on Artificial Intelligence, New York, NY, USA."},{"key":"ref_24","doi-asserted-by":"crossref","first-page":"7145","DOI":"10.1007\/s11042-020-10111-4","article-title":"DSFMA: Deeply supervised fully convolutional neural networks based on multi-level aggregation for saliency detection","volume":"80","author":"Ullah","year":"2021","journal-title":"Multimed. Tools Appl."},{"key":"ref_25","doi-asserted-by":"crossref","first-page":"7567","DOI":"10.1109\/TIP.2021.3106798","article-title":"Hierarchical Edge Refinement Network for Saliency Detection","volume":"30","author":"Song","year":"2021","journal-title":"IEEE Trans. Image Process."},{"key":"ref_26","doi-asserted-by":"crossref","first-page":"97","DOI":"10.1016\/0010-0285(80)90005-5","article-title":"A feature-integration theory of attention","volume":"12","author":"Treisman","year":"1980","journal-title":"Cogn. Psychol."},{"key":"ref_27","unstructured":"Simonyan, K., and Zisserman, A. (2014). Very deep convolutional networks for large-scale image recognition. arXiv."},{"key":"ref_28","doi-asserted-by":"crossref","first-page":"834","DOI":"10.1109\/TPAMI.2017.2699184","article-title":"Deeplab: Semantic image segmentation with deep convolutional nets, atrous convolution, and fully connected crfs","volume":"40","author":"Chen","year":"2017","journal-title":"IEEE Trans. Pattern Anal. Mach. Intell."},{"key":"ref_29","doi-asserted-by":"crossref","unstructured":"Woo, S., Park, J., Lee, J.Y., and Kweon, I.S. (2018, January 8\u201314). Cbam: Convolutional block attention module. Proceedings of the European Conference on Computer Vision (ECCV), Munich, Germany.","DOI":"10.1007\/978-3-030-01234-2_1"},{"key":"ref_30","doi-asserted-by":"crossref","unstructured":"Zhao, T., and Wu, X. (2019, January 15\u201320). Pyramid feature attention network for saliency detection. Proceedings of the IEEE\/CVF Conference on Computer Vision and Pattern Recognition, Long Beach, CA, USA.","DOI":"10.1109\/CVPR.2019.00320"},{"key":"ref_31","first-page":"898","article-title":"A Study on Various Attention for Improving Performance in Single Image Super Resolution","volume":"25","author":"Mun","year":"2020","journal-title":"J. Broadcast Eng."},{"key":"ref_32","unstructured":"Navalpakkam, V., and Itti, L. (2006, January 17\u201322). An integrated model of top-down and bottom-up attention for optimizing detection speed. Proceedings of the 2006 IEEE Computer Society Conference on Computer Vision and Pattern Recognition (CVPR\u201906), New York, NY, USA."},{"key":"ref_33","doi-asserted-by":"crossref","unstructured":"Garcia-Diaz, A., Fdez-Vidal, X.R., Pardo, X.M., and Dosil, R. (2009, January 18\u201321). Decorrelation and distinctiveness provide with human-like saliency. Proceedings of the International Conference on Advanced Concepts for Intelligent Vision Systems, Antwerp, Belgium.","DOI":"10.1007\/978-3-642-04697-1_32"},{"key":"ref_34","doi-asserted-by":"crossref","unstructured":"Zhang, L., Gu, Z., and Li, H. (2013, January 15\u201318). SDSP: A novel saliency detection method by combining simple priors. Proceedings of the 2013 IEEE International Conference on Image Processing, Melbourne, Australia.","DOI":"10.1109\/ICIP.2013.6738036"},{"key":"ref_35","first-page":"64","article-title":"Realistic avatar eye and head animation using a neurobiological model of visual attention","volume":"Volume 5200","author":"Itti","year":"2003","journal-title":"Proceedings of the Applications and Science of Neural Networks, Fuzzy Systems, and Evolutionary Computation VI"},{"key":"ref_36","doi-asserted-by":"crossref","first-page":"145","DOI":"10.1023\/A:1011139631724","article-title":"Modeling the shape of the scene: A holistic representation of the spatial envelope","volume":"42","author":"Oliva","year":"2001","journal-title":"Int. J. Comput. Vis."},{"key":"ref_37","doi-asserted-by":"crossref","first-page":"150","DOI":"10.1007\/s11263-010-0354-6","article-title":"Probabilistic multi-task learning for visual saliency estimation in video","volume":"90","author":"Li","year":"2010","journal-title":"Int. J. Comput. Vis."},{"key":"ref_38","unstructured":"Milanese, R. (1993). Detecting Salient Regions in an Image: From Biological Evidence to Computer Implementation. [Ph.D. Thesis, The University of Geneva]."},{"key":"ref_39","doi-asserted-by":"crossref","first-page":"64","DOI":"10.1016\/j.cviu.2004.09.005","article-title":"The emergence of attention by population-based inference and its role in distributed processing and cognitive control of vision","volume":"100","author":"Hamker","year":"2005","journal-title":"Comput. Vis. Image Underst."},{"key":"ref_40","doi-asserted-by":"crossref","first-page":"507","DOI":"10.1016\/0004-3702(95)00025-9","article-title":"Modeling visual attention via selective tuning","volume":"78","author":"Tsotsos","year":"1995","journal-title":"Artif. Intell."},{"key":"ref_41","doi-asserted-by":"crossref","unstructured":"Kootstra, G., Nederveen, A., and De Boer, B. (2008, January 1\u20134). Paying attention to symmetry. Proceedings of the British Machine Vision Conference (BMVC2008), The British Machine Vision Association and Society for Pattern Recognition, Leeds, UK.","DOI":"10.5244\/C.22.111"},{"key":"ref_42","doi-asserted-by":"crossref","first-page":"107","DOI":"10.1016\/S0042-6989(01)00250-4","article-title":"Modeling the role of salience in the allocation of overt visual attention","volume":"42","author":"Parkhurst","year":"2002","journal-title":"Vis. Res."},{"key":"ref_43","doi-asserted-by":"crossref","unstructured":"Deng, Z., Hu, X., Zhu, L., Xu, X., Qin, J., Han, G., and Heng, P.A. (2018, January 13\u201319). R3Net: Recurrent residual refinement network for saliency detection. Proceedings of the 27th International Joint Conference on Artificial Intelligence, Stockholm, Sweden.","DOI":"10.24963\/ijcai.2018\/95"},{"key":"ref_44","doi-asserted-by":"crossref","unstructured":"Hu, X., Zhu, L., Qin, J., Fu, C.W., and Heng, P.A. (2018, January 2\u20137). Recurrently aggregating deep features for salient object detection. Proceedings of the AAAI Conference on Artificial Intelligence, New Orleans, LA, USA.","DOI":"10.1609\/aaai.v32i1.12298"},{"key":"ref_45","doi-asserted-by":"crossref","unstructured":"Chen, S., Tan, X., Wang, B., and Hu, X. (2018, January 8\u201314). Reverse attention for salient object detection. Proceedings of the European Conference on Computer Vision (ECCV), Munich, Germany.","DOI":"10.1007\/978-3-030-01240-3_15"},{"key":"ref_46","doi-asserted-by":"crossref","unstructured":"Xie, S., and Tu, Z. (2015, January 7\u201313). Holistically-nested edge detection. Proceedings of the IEEE International Conference on Computer Vision, Washington, DC, USA.","DOI":"10.1109\/ICCV.2015.164"},{"key":"ref_47","doi-asserted-by":"crossref","unstructured":"Wu, Z., Su, L., and Huang, Q. (2019, January 15\u201320). Cascaded partial decoder for fast and accurate salient object detection. Proceedings of the IEEE\/CVF Conference on Computer Vision and Pattern Recognition, Long Beach, CA, USA.","DOI":"10.1109\/CVPR.2019.00403"},{"key":"ref_48","doi-asserted-by":"crossref","unstructured":"Ronneberger, O., Fischer, P., and Brox, T. (2015, January 5\u20139). U-net: Convolutional networks for biomedical image segmentation. Proceedings of the International Conference on Medical Image Computing and Computer-Assisted Intervention, Munich, Germany.","DOI":"10.1007\/978-3-319-24574-4_28"},{"key":"ref_49","doi-asserted-by":"crossref","unstructured":"Wang, T., Zhang, L., Wang, S., Lu, H., Yang, G., Ruan, X., and Borji, A. (2018, January 18\u201323). Detect globally, refine locally: A novel approach to saliency detection. Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition, Salt Lake City, UT, USA.","DOI":"10.1109\/CVPR.2018.00330"},{"key":"ref_50","doi-asserted-by":"crossref","unstructured":"Zhang, X., Wang, T., Qi, J., Lu, H., and Wang, G. (2018, January 18\u201323). Progressive attention guided recurrent network for salient object detection. Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition, Salt Lake City, UT, USA.","DOI":"10.1109\/CVPR.2018.00081"},{"key":"ref_51","doi-asserted-by":"crossref","unstructured":"Qin, X., Zhang, Z., Huang, C., Gao, C., Dehghan, M., and Jagersand, M. (2019, January 15\u201320). Basnet: Boundary-aware salient object detection. Proceedings of the IEEE\/CVF Conference on Computer Vision and Pattern Recognition, Long Beach, CA, USA.","DOI":"10.1109\/CVPR.2019.00766"},{"key":"ref_52","doi-asserted-by":"crossref","unstructured":"Chen, Z., Xu, Q., Cong, R., and Huang, Q. (2020, January 7\u201312). Global context-aware progressive aggregation network for salient object detection. Proceedings of the AAAI Conference on Artificial Intelligence, New York, NY, USA.","DOI":"10.1609\/aaai.v34i07.6633"},{"key":"ref_53","doi-asserted-by":"crossref","first-page":"2274","DOI":"10.1109\/TIP.2017.2682981","article-title":"RGBD salient object detection via deep fusion","volume":"26","author":"Qu","year":"2017","journal-title":"IEEE Trans. Image Process."},{"key":"ref_54","doi-asserted-by":"crossref","first-page":"3171","DOI":"10.1109\/TCYB.2017.2761775","article-title":"CNNs-based RGB-D saliency detection via cross-view transfer and multiview fusion","volume":"48","author":"Han","year":"2017","journal-title":"IEEE Trans. Cybern."},{"key":"ref_55","doi-asserted-by":"crossref","unstructured":"Piao, Y., Ji, W., Li, J., Zhang, M., and Lu, H. (2019, January 27\u201328). Depth-induced multi-scale recurrent attention network for saliency detection. Proceedings of the IEEE\/CVF International Conference on Computer Vision, Seoul, Korea.","DOI":"10.1109\/ICCV.2019.00735"},{"key":"ref_56","doi-asserted-by":"crossref","first-page":"376","DOI":"10.1016\/j.patcog.2018.08.007","article-title":"Multi-modal fusion network with multi-scale multi-path and cross-modal interactions for RGB-D salient object detection","volume":"86","author":"Chen","year":"2019","journal-title":"Pattern Recognit."},{"key":"ref_57","unstructured":"Recommendation, ITURBT (2015). 709-6: Parameter Values for the HDTV Standards for Production and International Programme Exchange, ITU. Basic parameter values for the HDTV standard for the studio and for international programme exchange, now ITU-R BT."},{"key":"ref_58","unstructured":"Munsell, A.H. (1907). A Color Notation, GH Ellis Company."},{"key":"ref_59","doi-asserted-by":"crossref","first-page":"236","DOI":"10.2307\/1412843","article-title":"A pigment color system and notation","volume":"23","author":"Munsell","year":"1912","journal-title":"Am. J. Psychol."},{"key":"ref_60","doi-asserted-by":"crossref","unstructured":"Hu, J., Shen, L., and Sun, G. (2018, January 18\u201323). Squeeze-and-excitation networks. Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition, Salt Lake City, UT, USA.","DOI":"10.1109\/CVPR.2018.00745"},{"key":"ref_61","doi-asserted-by":"crossref","unstructured":"Peng, C., Zhang, X., Yu, G., Luo, G., and Sun, J. (2017, January 21\u201326). Large kernel matters\u2013improve semantic segmentation by global convolutional network. Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition, Honolulu, HI, USA.","DOI":"10.1109\/CVPR.2017.189"},{"key":"ref_62","doi-asserted-by":"crossref","unstructured":"Chen, L., Zhang, H., Xiao, J., Nie, L., Shao, J., Liu, W., and Chua, T.S. (2017, January 21\u201326). Sca-cnn: Spatial and channel-wise attention in convolutional networks for image captioning. Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition, Honolulu, HI, USA.","DOI":"10.1109\/CVPR.2017.667"},{"key":"ref_63","doi-asserted-by":"crossref","first-page":"107303","DOI":"10.1016\/j.patcog.2020.107303","article-title":"CAGNet: Content-aware guidance for salient object detection","volume":"103","author":"Mohammadi","year":"2020","journal-title":"Pattern Recognit."},{"key":"ref_64","doi-asserted-by":"crossref","unstructured":"He, K., Zhang, X., Ren, S., and Sun, J. (2016, January 27\u201330). Deep residual learning for image recognition. Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition, Las Vegas, NV, USA.","DOI":"10.1109\/CVPR.2016.90"},{"key":"ref_65","doi-asserted-by":"crossref","unstructured":"He, K., Zhang, X., Ren, S., and Sun, J. (2016, January 11\u201314). Identity mappings in deep residual networks. Proceedings of the European Conference on Computer Vision, Amsterdam, The Netherlands.","DOI":"10.1007\/978-3-319-46493-0_38"},{"key":"ref_66","doi-asserted-by":"crossref","unstructured":"Yang, C., Zhang, L., Lu, H., Ruan, X., and Yang, M.H. (2013, January 23\u201328). Saliency detection via graph-based manifold ranking. Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition, Portland, OR, USA.","DOI":"10.1109\/CVPR.2013.407"},{"key":"ref_67","doi-asserted-by":"crossref","unstructured":"Wang, L., Lu, H., Wang, Y., Feng, M., Wang, D., Yin, B., and Ruan, X. (2017, January 21\u201326). Learning to detect salient objects with image-level supervision. Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition, Honolulu, HI, USA.","DOI":"10.1109\/CVPR.2017.404"},{"key":"ref_68","doi-asserted-by":"crossref","unstructured":"Deng, J., Dong, W., Socher, R., Li, L.J., Li, K., and Fei-Fei, L. (2009, January 20\u201325). Imagenet: A large-scale hierarchical image database. Proceedings of the 2009 IEEE Conference on Computer Vision and Pattern Recognition, Miami, FL, USA.","DOI":"10.1109\/CVPR.2009.5206848"},{"key":"ref_69","doi-asserted-by":"crossref","unstructured":"Xiao, J., Hays, J., Ehinger, K.A., Oliva, A., and Torralba, A. (2010, January 13\u201318). Sun database: Large-scale scene recognition from abbey to zoo. Proceedings of the 2010 IEEE Computer Society Conference on Computer Vision and Pattern Recognition, San Francisco, CA, USA.","DOI":"10.1109\/CVPR.2010.5539970"},{"key":"ref_70","doi-asserted-by":"crossref","first-page":"717","DOI":"10.1109\/TPAMI.2015.2465960","article-title":"Hierarchical image saliency detection on extended CSSD","volume":"38","author":"Shi","year":"2015","journal-title":"IEEE Trans. Pattern Anal. Mach. Intell."},{"key":"ref_71","unstructured":"Li, G., and Yu, Y. (2015, January 7\u201312). Visual saliency based on multiscale deep features. Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition, Boston, MA, USA."},{"key":"ref_72","doi-asserted-by":"crossref","first-page":"303","DOI":"10.1007\/s11263-009-0275-4","article-title":"The pascal visual object classes (voc) challenge","volume":"88","author":"Everingham","year":"2010","journal-title":"Int. J. Comput. Vis."},{"key":"ref_73","doi-asserted-by":"crossref","unstructured":"Li, Y., Hou, X., Koch, C., Rehg, J.M., and Yuille, A.L. (2014, January 23\u201328). The secrets of salient object segmentation. Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition, Columbus, OH, USA.","DOI":"10.1109\/CVPR.2014.43"},{"key":"ref_74","doi-asserted-by":"crossref","unstructured":"Zhao, R., Ouyang, W., Li, H., and Wang, X. (2015, January 7\u201312). Saliency detection by multi-context deep learning. Proceedings of the Saliency Detection by Multi-Context Deep Learning, Boston, MA, USA.","DOI":"10.1109\/CVPR.2015.7298731"},{"key":"ref_75","unstructured":"Kingma, D.P., and Ba, J. (2014). Adam: A method for stochastic optimization. arXiv."}],"container-title":["Sensors"],"original-title":[],"language":"en","link":[{"URL":"https:\/\/www.mdpi.com\/1424-8220\/22\/9\/3588\/pdf","content-type":"unspecified","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2025,10,10]],"date-time":"2025-10-10T23:08:09Z","timestamp":1760137689000},"score":1,"resource":{"primary":{"URL":"https:\/\/www.mdpi.com\/1424-8220\/22\/9\/3588"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2022,5,9]]},"references-count":75,"journal-issue":{"issue":"9","published-online":{"date-parts":[[2022,5]]}},"alternative-id":["s22093588"],"URL":"https:\/\/doi.org\/10.3390\/s22093588","relation":{},"ISSN":["1424-8220"],"issn-type":[{"value":"1424-8220","type":"electronic"}],"subject":[],"published":{"date-parts":[[2022,5,9]]}}}