{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2026,1,18]],"date-time":"2026-01-18T06:49:00Z","timestamp":1768718940675,"version":"3.49.0"},"reference-count":39,"publisher":"MDPI AG","issue":"3","license":[{"start":{"date-parts":[[2023,1,17]],"date-time":"2023-01-17T00:00:00Z","timestamp":1673913600000},"content-version":"vor","delay-in-days":0,"URL":"https:\/\/creativecommons.org\/licenses\/by\/4.0\/"}],"funder":[{"name":"Scientific Research Fund of the Zhejiang Provincial Education Department","award":["Y202147449"],"award-info":[{"award-number":["Y202147449"]}]},{"name":"Scientific Research Fund of the Zhejiang Provincial Education Department","award":["LGG20F010010"],"award-info":[{"award-number":["LGG20F010010"]}]},{"name":"Scientific Research Fund of the Zhejiang Provincial Education Department","award":["2020AY10024"],"award-info":[{"award-number":["2020AY10024"]}]},{"name":"Scientific Research Fund of the Zhejiang Provincial Education Department","award":["2022AD10017"],"award-info":[{"award-number":["2022AD10017"]}]},{"name":"Zhejiang Public Welfare Technology Research Project Fund of China","award":["Y202147449"],"award-info":[{"award-number":["Y202147449"]}]},{"name":"Zhejiang Public Welfare Technology Research Project Fund of China","award":["LGG20F010010"],"award-info":[{"award-number":["LGG20F010010"]}]},{"name":"Zhejiang Public Welfare Technology Research Project Fund of China","award":["2020AY10024"],"award-info":[{"award-number":["2020AY10024"]}]},{"name":"Zhejiang Public Welfare Technology Research Project Fund of China","award":["2022AD10017"],"award-info":[{"award-number":["2022AD10017"]}]},{"name":"Jiaxing Science and Technology Project","award":["Y202147449"],"award-info":[{"award-number":["Y202147449"]}]},{"name":"Jiaxing Science and Technology Project","award":["LGG20F010010"],"award-info":[{"award-number":["LGG20F010010"]}]},{"name":"Jiaxing Science and Technology Project","award":["2020AY10024"],"award-info":[{"award-number":["2020AY10024"]}]},{"name":"Jiaxing Science and Technology Project","award":["2022AD10017"],"award-info":[{"award-number":["2022AD10017"]}]}],"content-domain":{"domain":[],"crossmark-restriction":false},"short-container-title":["Sensors"],"abstract":"<jats:p>Detecting irregular or arbitrary shape text in natural scene images is a challenging task that has recently attracted considerable attention from research communities. However, limited by the CNN receptive field, these methods cannot directly capture relations between distant component regions by local convolutional operators. In this paper, we propose a novel method that can effectively and robustly detect irregular text in natural scene images. First, we employ a fully convolutional network architecture based on VGG16_BN to generate text components via the estimated character center points, which can ensure a high text component detection recall rate and fewer noncharacter text components. Second, text line grouping is treated as a problem of inferring the adjacency relations of text components with a graph convolution network (GCN). Finally, to evaluate our algorithm, we compare it with other existing algorithms by performing experiments on three public datasets: ICDAR2013, CTW-1500 and MSRA-TD500. The results show that the proposed method handles irregular scene text well and that it achieves promising results on these three public datasets.<\/jats:p>","DOI":"10.3390\/s23031070","type":"journal-article","created":{"date-parts":[[2023,1,17]],"date-time":"2023-01-17T05:40:02Z","timestamp":1673934002000},"page":"1070","update-policy":"https:\/\/doi.org\/10.3390\/mdpi_crossmark_policy","source":"Crossref","is-referenced-by-count":10,"title":["Irregular Scene Text Detection Based on a Graph Convolutional Network"],"prefix":"10.3390","volume":"23","author":[{"given":"Shiyu","family":"Zhang","sequence":"first","affiliation":[{"name":"College of Science, Jiangxi University of Science and Technology, Ganzhou 341000, China"},{"name":"College of Information Science and Engineering, Jiaxing University, Jiaxing 314001, China"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Caiying","family":"Zhou","sequence":"additional","affiliation":[{"name":"College of Science, Jiangxi University of Science and Technology, Ganzhou 341000, China"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Yonggang","family":"Li","sequence":"additional","affiliation":[{"name":"College of Information Science and Engineering, Jiaxing University, Jiaxing 314001, China"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Xianchao","family":"Zhang","sequence":"additional","affiliation":[{"name":"Key Laboratory of Medical Electronics and Digital Health of Zhejiang Province, Jiaxing University, Jiaxing 314001, China"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"ORCID":"https:\/\/orcid.org\/0000-0002-5385-0485","authenticated-orcid":false,"given":"Lihua","family":"Ye","sequence":"additional","affiliation":[{"name":"College of Information Science and Engineering, Jiaxing University, Jiaxing 314001, China"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"ORCID":"https:\/\/orcid.org\/0000-0002-3740-8101","authenticated-orcid":false,"given":"Yuanwang","family":"Wei","sequence":"additional","affiliation":[{"name":"College of Information Science and Engineering, Jiaxing University, Jiaxing 314001, China"},{"name":"Key Laboratory of Medical Electronics and Digital Health of Zhejiang Province, Jiaxing University, Jiaxing 314001, China"}],"role":[{"role":"author","vocabulary":"crossref"}]}],"member":"1968","published-online":{"date-parts":[[2023,1,17]]},"reference":[{"key":"ref_1","doi-asserted-by":"crossref","first-page":"161","DOI":"10.1007\/s11263-020-01369-0","article-title":"Scene text detection and recognition: The deep learning era","volume":"129","author":"Long","year":"2021","journal-title":"Int. J. Comput. Vis."},{"key":"ref_2","doi-asserted-by":"crossref","first-page":"2864","DOI":"10.1109\/TIP.2022.3141844","article-title":"Cm-net: Concentric mask based arbitrary-shaped text detection","volume":"31","author":"Yang","year":"2022","journal-title":"IEEE Trans. Image Process."},{"key":"ref_3","doi-asserted-by":"crossref","first-page":"8635","DOI":"10.1109\/TCSVT.2022.3194835","article-title":"Robust Scene Text Detection for Partially Annotated Training Data","volume":"32","author":"Keserwani","year":"2022","journal-title":"IEEE Trans. Circuits Syst. Video Technol."},{"key":"ref_4","unstructured":"Tian, Z., Huang, W., He, T., He, P., and Qiao, Y. (2016). Proceedings of the European Conference on Computer Vision, Springer."},{"key":"ref_5","doi-asserted-by":"crossref","unstructured":"Zhu, Y., Chen, J., Liang, L., Kuang, Z., and Zhang, W. (2021, January 20\u201325). Fourier Contour Embedding for Arbitrary-Shaped Text Detection. Proceedings of the IEEE\/CVF Conference on Computer Vision and Pattern Recognition, Nashville, TN, USA.","DOI":"10.1109\/CVPR46437.2021.00314"},{"key":"ref_6","doi-asserted-by":"crossref","first-page":"36802","DOI":"10.1109\/ACCESS.2021.3063030","article-title":"Quadbox: Quadrilateral Bounding Box Based Scene Text Detection Using Vector Regression","volume":"9","author":"Keserwani","year":"2021","journal-title":"IEEE Access"},{"key":"ref_7","doi-asserted-by":"crossref","unstructured":"Long, S., Ruan, J., Zhang, W., He, X., Wu, W., and Yao, C. (2018, January 8\u201314). TextSnake: A Flexible Representation for Detecting Text of Arbitrary Shapes. Proceedings of the European Conference on Computer Vision, Munich, Germany.","DOI":"10.1007\/978-3-030-01216-8_2"},{"key":"ref_8","doi-asserted-by":"crossref","unstructured":"Zhang, S.X., Zhu, X., Hou, J.B., Liu, C., Yang, C., Wang, H., and Yin, X.C. (2020, January 13\u201319). Deep relational reasoning graph network for arbitrary shape text detection. Proceedings of the IEEE\/CVF Conference on Computer Vision and Pattern Recognition, Seattle, WA, USA.","DOI":"10.1109\/CVPR42600.2020.00972"},{"key":"ref_9","first-page":"1","article-title":"Text recognition in the wild: A survey","volume":"54","author":"Chen","year":"2021","journal-title":"ACM Comput. Surv. (CSUR)"},{"key":"ref_10","doi-asserted-by":"crossref","first-page":"89","DOI":"10.1016\/j.image.2018.02.016","article-title":"Multi-oriented text detection from natural scene images based on a CNN and pruning non-adjacent graph edges","volume":"64","author":"Wei","year":"2018","journal-title":"Signal Process. Image Commun."},{"key":"ref_11","doi-asserted-by":"crossref","unstructured":"Li, X., Wang, W., Hou, W., Liu, R.Z., Lu, T., and Yang, J. (2018). Shape Robust Text Detection with Progressive Scale Expansion Network. arXiv.","DOI":"10.1109\/CVPR.2019.00956"},{"key":"ref_12","unstructured":"Dan, D., Liu, H., Li, X., and Deng, C. (2018). PixelLink: Detecting Scene Text via Instance Segmentation. arXiv."},{"key":"ref_13","doi-asserted-by":"crossref","unstructured":"Baek, Y., Lee, B., Han, D., Yun, S., and Lee, H. (2019, January 15\u201320). Character Region Awareness for Text Detection. Proceedings of the 2019 IEEE\/CVF Conference on Computer Vision and Pattern Recognition (CVPR), Long Beach, CA, USA.","DOI":"10.1109\/CVPR.2019.00959"},{"key":"ref_14","doi-asserted-by":"crossref","unstructured":"Liu, Y., Chen, H., Shen, C., He, T., Jin, L., and Wang, L. (2020, January 13\u201319). Abcnet: Real-time scene text spotting with adaptive bezier-curve network. Proceedings of the IEEE\/CVF Conference on Computer Vision and Pattern Recognition, Seattle, WA, USA.","DOI":"10.1109\/CVPR42600.2020.00983"},{"key":"ref_15","doi-asserted-by":"crossref","unstructured":"Wang, Y., Mamat, H., Xu, X., Aysa, A., and Ubul, K. (2022). Scene Uyghur Text Detection Based on Fine-Grained Feature Representation. Sensors, 22.","DOI":"10.3390\/s22124372"},{"key":"ref_16","doi-asserted-by":"crossref","unstructured":"Zhang, S.X., Zhu, X., Yang, C., Wang, H., and Yin, X.C. (2021, January 19\u201325). Adaptive boundary proposal network for arbitrary shape text detection. Proceedings of the IEEE\/CVF International Conference on Computer Vision, Virtual Conference.","DOI":"10.1109\/ICCV48922.2021.00134"},{"key":"ref_17","doi-asserted-by":"crossref","unstructured":"Liao, M., Shi, B., Bai, X., Wang, X., and Liu, W. (2016). TextBoxes: A Fast Text Detector with a Single Deep Neural Network. arXiv.","DOI":"10.1609\/aaai.v31i1.11196"},{"key":"ref_18","doi-asserted-by":"crossref","unstructured":"Zhou, X., Yao, C., Wen, H., Wang, Y., Zhou, S., He, W., and Liang, J. (2017, January 21\u201326). EAST: An Efficient and Accurate Scene Text Detector. Proceedings of the 2017 IEEE Conference on Computer Vision and Pattern Recognition (CVPR), Honolulu, HI, USA.","DOI":"10.1109\/CVPR.2017.283"},{"key":"ref_19","doi-asserted-by":"crossref","unstructured":"Tian, Z., Shu, M., Lyu, P., Li, R., Zhou, C., Shen, X., and Jia, J. (2019, January 15\u201320). Learning Shape-Aware Embedding for Scene Text Detection. Proceedings of the 2019 IEEE\/CVF Conference on Computer Vision and Pattern Recognition (CVPR), Long Beach, CA, USA.","DOI":"10.1109\/CVPR.2019.00436"},{"key":"ref_20","doi-asserted-by":"crossref","unstructured":"Zhang, S.X., Zhu, X., Hou, J.B., Yang, C., and Yin, X.C. (IEEE Trans. Neural Netw. Learn. Syst., 2022). Kernel Proposal Network for Arbitrary Shape Text Detection, IEEE Trans. Neural Netw. Learn. Syst., Early Access.","DOI":"10.1109\/TNNLS.2022.3152596"},{"key":"ref_21","doi-asserted-by":"crossref","first-page":"31","DOI":"10.1007\/s10032-019-00334-z","article-title":"Total-text: Toward orientation robustness in scene text detection","volume":"23","author":"Chan","year":"2020","journal-title":"Int. J. Doc. Anal. Recognit. (IJDAR)"},{"key":"ref_22","first-page":"1","article-title":"Faster r-cnn: Towards real-time object detection with region proposal networks","volume":"28","author":"Ren","year":"2015","journal-title":"Adv. Neural Inf. Process. Syst."},{"key":"ref_23","unstructured":"Shrivastava, A., Gupta, A., and Girshick, R. (July, January 26). Training region-based object detectors with online hard example mining. Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition, Las Vegas, NV, USA."},{"key":"ref_24","first-page":"1","article-title":"Attention is all you need","volume":"30","author":"Vaswani","year":"2017","journal-title":"Adv. Neural Inf. Process. Syst."},{"key":"ref_25","doi-asserted-by":"crossref","unstructured":"Gu, J., Hu, H., Wang, L., Wei, Y., and Dai, J. (2018, January 8\u201314). Learning region features for object detection. Proceedings of the European Conference on Computer Vision (ECCV), Munich, Germany.","DOI":"10.1007\/978-3-030-01258-8_24"},{"key":"ref_26","doi-asserted-by":"crossref","first-page":"280","DOI":"10.1007\/s10032-006-0014-0","article-title":"Object count\/area graphs for the evaluation of object detection and segmentation algorithms","volume":"8","author":"Wolf","year":"2006","journal-title":"Int. J. Doc. Anal. Recognit."},{"key":"ref_27","doi-asserted-by":"crossref","first-page":"84","DOI":"10.1145\/3065386","article-title":"Imagenet classification with deep convolutional neural networks","volume":"60","author":"Krizhevsky","year":"2017","journal-title":"Commun. ACM"},{"key":"ref_28","doi-asserted-by":"crossref","first-page":"1","DOI":"10.1016\/j.image.2016.10.003","article-title":"Text detection in scene images based on exhaustive segmentation","volume":"50","author":"Wei","year":"2017","journal-title":"Signal Process. Image Commun. Publ. Eur. Assoc. Signal Process."},{"key":"ref_29","doi-asserted-by":"crossref","first-page":"96424","DOI":"10.1109\/ACCESS.2019.2929819","article-title":"Convolutional Regression Network for Multi-oriented Text Detection","volume":"7","author":"Gao","year":"2019","journal-title":"IEEE Access"},{"key":"ref_30","doi-asserted-by":"crossref","unstructured":"Jeon, M., and Jeong, Y.S. (2020). Compact and accurate scene text detector. Appl. Sci., 10.","DOI":"10.3390\/app10062096"},{"key":"ref_31","doi-asserted-by":"crossref","first-page":"106954","DOI":"10.1016\/j.patcog.2019.06.020","article-title":"Detecting Dense and Arbitrary-shaped Scene Text by Instance-aware Component Grouping","volume":"96","author":"Tang","year":"2019","journal-title":"Pattern Recognit."},{"key":"ref_32","doi-asserted-by":"crossref","first-page":"5566","DOI":"10.1109\/TIP.2019.2900589","article-title":"TextField: Learning A Deep Direction Field for Irregular Scene Text Detection","volume":"28","author":"Xu","year":"2019","journal-title":"IEEE Trans. Image Process."},{"key":"ref_33","doi-asserted-by":"crossref","unstructured":"Wan, Q., Ji, H., and Shen, L. (2021, January 20\u201325). Self-attention based text knowledge mining for text detection. Proceedings of the IEEE\/CVF Conference on Computer Vision and Pattern Recognition, Nashville, TN, USA.","DOI":"10.1109\/CVPR46437.2021.00592"},{"key":"ref_34","doi-asserted-by":"crossref","unstructured":"Wang, X., Jiang, Y., Luo, Z., Liu, C.L., Choi, H., and Kim, S. (2019, January 15\u201320). Arbitrary Shape Scene Text Detection with Adaptive Text Region Representation. Proceedings of the IEEE\/CVF Conference on Computer Vision and Pattern Recognition, Long Beach, CA, USA.","DOI":"10.1109\/CVPR.2019.00661"},{"key":"ref_35","doi-asserted-by":"crossref","unstructured":"Wang, X., Zheng, S., Zhang, C., Li, R., and Gui, L. (2021). R-YOLO: A real-time text detector for natural scenes with arbitrary rotation. Sensors, 21.","DOI":"10.3390\/s21030888"},{"key":"ref_36","doi-asserted-by":"crossref","unstructured":"Raisi, Z., Naiel, M.A., Younes, G., Wardell, S., and Zelek, J.S. (2021, January 20\u201325). Transformer-based text detection in the wild. Proceedings of the IEEE\/CVF Conference on Computer Vision and Pattern Recognition, Nashville, TN, USA.","DOI":"10.1109\/CVPRW53098.2021.00353"},{"key":"ref_37","doi-asserted-by":"crossref","first-page":"919","DOI":"10.1109\/TPAMI.2022.3155612","article-title":"Real-time scene text detection with differentiable binarization and adaptive scale fusion","volume":"45","author":"Liao","year":"2022","journal-title":"IEEE Trans. Pattern Anal. Mach. Intell."},{"key":"ref_38","doi-asserted-by":"crossref","unstructured":"Feng, W., He, W., Yin, F., Zhang, X.Y., and Liu, C.L. (2020, January 13\u201319). TextDragon: An End-to-End Framework for Arbitrary Shaped Text Spotting. Proceedings of the 2019 IEEE\/CVF International Conference on Computer Vision (ICCV), Seattle, WA, USA.","DOI":"10.1109\/ICCV.2019.00917"},{"key":"ref_39","doi-asserted-by":"crossref","unstructured":"Wang, Y., Xie, H., Zha, Z., Xing, M., Fu, Z., and Zhang, Y. (2020, January 13\u201319). ContourNet: Taking a Further Step toward Accurate Arbitrary-shaped Scene Text Detection. Proceedings of the 2020 IEEE\/CVF Conference on Computer Vision and Pattern Recognition (CVPR), Seattle, WA, USA.","DOI":"10.1109\/CVPR42600.2020.01177"}],"container-title":["Sensors"],"original-title":[],"language":"en","link":[{"URL":"https:\/\/www.mdpi.com\/1424-8220\/23\/3\/1070\/pdf","content-type":"unspecified","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2025,10,10]],"date-time":"2025-10-10T18:08:18Z","timestamp":1760119698000},"score":1,"resource":{"primary":{"URL":"https:\/\/www.mdpi.com\/1424-8220\/23\/3\/1070"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2023,1,17]]},"references-count":39,"journal-issue":{"issue":"3","published-online":{"date-parts":[[2023,2]]}},"alternative-id":["s23031070"],"URL":"https:\/\/doi.org\/10.3390\/s23031070","relation":{},"ISSN":["1424-8220"],"issn-type":[{"value":"1424-8220","type":"electronic"}],"subject":[],"published":{"date-parts":[[2023,1,17]]}}}