{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2025,10,12]],"date-time":"2025-10-12T01:45:27Z","timestamp":1760233527070,"version":"build-2065373602"},"reference-count":20,"publisher":"MDPI AG","issue":"2","license":[{"start":{"date-parts":[[2021,1,28]],"date-time":"2021-01-28T00:00:00Z","timestamp":1611792000000},"content-version":"vor","delay-in-days":0,"URL":"https:\/\/creativecommons.org\/licenses\/by\/4.0\/"}],"content-domain":{"domain":[],"crossmark-restriction":false},"short-container-title":["Symmetry"],"abstract":"<jats:p>With continuous developments in deep learning, image semantic segmentation technology has also undergone great advancements and been widely used in many fields with higher segmentation accuracy. This paper proposes an image semantic segmentation algorithm based on a deep neural network. Based on the Mask Scoring R-CNN, this algorithm uses a symmetrical feature pyramid network and adds a multiple-threshold architecture to improve the sample screening precision. We employ a probability model to optimize the mask branch of the model further to improve the algorithm accuracy for the segmentation of image edges. In addition, we adjust the loss function so that the experimental effect can be optimized. The experiments reveal that the algorithm improves the results.<\/jats:p>","DOI":"10.3390\/sym13020207","type":"journal-article","created":{"date-parts":[[2021,1,28]],"date-time":"2021-01-28T01:22:23Z","timestamp":1611796943000},"page":"207","update-policy":"https:\/\/doi.org\/10.3390\/mdpi_crossmark_policy","source":"Crossref","is-referenced-by-count":7,"title":["Image Semantic Segmentation Use Multiple-Threshold Probabilistic R-CNN with Feature Fusion"],"prefix":"10.3390","volume":"13","author":[{"ORCID":"https:\/\/orcid.org\/0000-0002-3048-6000","authenticated-orcid":false,"given":"Jianxin","family":"Liu","sequence":"first","affiliation":[{"name":"School of Computer Science and Technology, Qilu University of Technology (Shandong Academy of Sciences), Jinan 250353, China"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Yushui","family":"Geng","sequence":"additional","affiliation":[{"name":"School of Computer Science and Technology, Qilu University of Technology (Shandong Academy of Sciences), Jinan 250353, China"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Jing","family":"Zhao","sequence":"additional","affiliation":[{"name":"School of Computer Science and Technology, Qilu University of Technology (Shandong Academy of Sciences), Jinan 250353, China"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"ORCID":"https:\/\/orcid.org\/0000-0001-7148-6106","authenticated-orcid":false,"given":"Kang","family":"Zhang","sequence":"additional","affiliation":[{"name":"School of Computer Science and Technology, Qilu University of Technology (Shandong Academy of Sciences), Jinan 250353, China"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Wenxiao","family":"Li","sequence":"additional","affiliation":[{"name":"School of Computer Science and Technology, Qilu University of Technology (Shandong Academy of Sciences), Jinan 250353, China"}],"role":[{"role":"author","vocabulary":"crossref"}]}],"member":"1968","published-online":{"date-parts":[[2021,1,28]]},"reference":[{"key":"ref_1","unstructured":"Simonyan, K., and Zisserman, A. (2014). Very deep convolutional networks for large-scale image recognition. arXiv."},{"key":"ref_2","doi-asserted-by":"crossref","unstructured":"He, K., Zhang, X., Ren, S., and Sun, J. (2016, January 27\u201330). Deep residual learning for image recognition. Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition, Las Vegas, NV, USA.","DOI":"10.1109\/CVPR.2016.90"},{"key":"ref_3","doi-asserted-by":"crossref","unstructured":"Long, J., Shelhamer, E., and Darrell, T. (2015, January 7\u201312). Fully convolutional networks for semantic segmentation. Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition, Boston, MA, USA.","DOI":"10.1109\/CVPR.2015.7298965"},{"key":"ref_4","unstructured":"Pinheiro, P.O., Collobert, R., and Doll\u00e1r, P. (2015). Learning to segment object candidates. arXiv."},{"key":"ref_5","doi-asserted-by":"crossref","unstructured":"He, K., Gkioxari, G., Doll\u00e1r, P., and Girshick, R. (2017, January 22\u201329). Mask r-cnn. Proceedings of the IEEE International Conference on Computer Vision, Venice, Italy.","DOI":"10.1109\/ICCV.2017.322"},{"key":"ref_6","doi-asserted-by":"crossref","unstructured":"Huang, Z., Huang, L., Gong, Y., Huang, C., and Wang, X. (2019, January 16\u201320). Mask scoring r-cnn. Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition, Long Beach, CA, USA.","DOI":"10.1109\/CVPR.2019.00657"},{"key":"ref_7","doi-asserted-by":"crossref","first-page":"264","DOI":"10.1016\/j.biosystemseng.2020.03.008","article-title":"Instance segmentation of apple flowers using the improved mask R\u2013CNN model","volume":"193","author":"Tian","year":"2020","journal-title":"Biosyst. Eng."},{"key":"ref_8","doi-asserted-by":"crossref","unstructured":"Zhang, B., and Zhang, J. (2020). A Traffic Surveillance System for Obtaining Comprehensive Information of the Passing Vehicles Based on Instance Segmentation. IEEE Trans. Intell. Transp. Syst.","DOI":"10.1109\/TITS.2020.3001154"},{"key":"ref_9","unstructured":"Liu, D., Zhang, D., Song, Y., Huang, H., and Cai, W. (2020). Cell r-cnn v3: A novel panoptic paradigm for instance segmentation in biomedical images. arXiv."},{"key":"ref_10","doi-asserted-by":"crossref","unstructured":"Lin, T.Y., Maire, M., Belongie, S., Hays, J., Perona, P., Ramanan, D., Doll\u00e1r, P., and Zitnick, C.L. (2014, January 6\u201312). Microsoft coco: Common objects in context. Proceedings of the European Conference on Computer Vision, Zurich, Switzerland.","DOI":"10.1007\/978-3-319-10602-1_48"},{"key":"ref_11","doi-asserted-by":"crossref","unstructured":"Cai, Z., and Vasconcelos, N. (2019). Cascade R-CNN: High quality object detection and instance segmentation. IEEE Trans. Pattern Anal. Mach. Intell.","DOI":"10.1109\/CVPR.2018.00644"},{"key":"ref_12","doi-asserted-by":"crossref","unstructured":"Lin, T.Y., Doll\u00e1r, P., Girshick, R., He, K., Hariharan, B., and Belongie, S. (2017, January 21\u201326). Feature pyramid networks for object detection. Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition, Honolulu, HI, USA.","DOI":"10.1109\/CVPR.2017.106"},{"key":"ref_13","doi-asserted-by":"crossref","unstructured":"Wang, X., Girshick, R., Gupta, A., and He, K. (2018, January 18\u201322). Non-local neural networks. Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition, Salt Lake City, UT, USA.","DOI":"10.1109\/CVPR.2018.00813"},{"key":"ref_14","unstructured":"Ren, S., He, K., Girshick, R., and Sun, J. (2015). Faster r-cnn: Towards real-time object detection with region proposal networks. arXiv."},{"key":"ref_15","doi-asserted-by":"crossref","unstructured":"Girshick, R. (2015, January 7\u201313). Fast r-cnn. Proceedings of the IEEE International Conference on Computer Vision, Santiago, Chile.","DOI":"10.1109\/ICCV.2015.169"},{"key":"ref_16","unstructured":"Kr\u00e4henb\u00fchl, P., and Koltun, V. (2011). Efficient Inference in Fully Connected Crfs with Gaussian Edge Potentials. arXiv."},{"key":"ref_17","doi-asserted-by":"crossref","unstructured":"Li, Y., Qi, H., Dai, J., Ji, X., and Wei, Y. (2017, January 21\u201326). Fully convolutional instance-aware semantic segmentation. Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition, Honolulu, HI, USA.","DOI":"10.1109\/CVPR.2017.472"},{"key":"ref_18","doi-asserted-by":"crossref","unstructured":"Chen, L.C., Hermans, A., Papandreou, G., Schroff, F., Wang, P., and Adam, H. (2018, January 18\u201322). Masklab: Instance segmentation by refining object detection with semantic and direction features. Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition, Salt Lake City, UT, USA.","DOI":"10.1109\/CVPR.2018.00422"},{"key":"ref_19","doi-asserted-by":"crossref","first-page":"1983","DOI":"10.1007\/s11554-020-01007-5","article-title":"Joint multi-task cascade for instance segmentation","volume":"17","author":"Wen","year":"2020","journal-title":"J. Real-Time Image Process."},{"key":"ref_20","doi-asserted-by":"crossref","unstructured":"Cao, J., Anwer, R.M., Cholakkal, H., Khan, F.S., Pang, Y., and Shao, L. (2020). SipMask: Spatial Information Preservation for Fast Image and Video Instance Segmentation. arXiv.","DOI":"10.1007\/978-3-030-58568-6_1"}],"container-title":["Symmetry"],"original-title":[],"language":"en","link":[{"URL":"https:\/\/www.mdpi.com\/2073-8994\/13\/2\/207\/pdf","content-type":"unspecified","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2025,10,11]],"date-time":"2025-10-11T05:16:23Z","timestamp":1760159783000},"score":1,"resource":{"primary":{"URL":"https:\/\/www.mdpi.com\/2073-8994\/13\/2\/207"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2021,1,28]]},"references-count":20,"journal-issue":{"issue":"2","published-online":{"date-parts":[[2021,2]]}},"alternative-id":["sym13020207"],"URL":"https:\/\/doi.org\/10.3390\/sym13020207","relation":{},"ISSN":["2073-8994"],"issn-type":[{"type":"electronic","value":"2073-8994"}],"subject":[],"published":{"date-parts":[[2021,1,28]]}}}