{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2025,11,5]],"date-time":"2025-11-05T11:24:36Z","timestamp":1762341876768,"version":"build-2065373602"},"reference-count":33,"publisher":"MDPI AG","issue":"12","license":[{"start":{"date-parts":[[2022,6,9]],"date-time":"2022-06-09T00:00:00Z","timestamp":1654732800000},"content-version":"vor","delay-in-days":0,"URL":"https:\/\/creativecommons.org\/licenses\/by\/4.0\/"}],"funder":[{"name":"National Science Foundation of China (NSFC)","award":["61862061","62061045","61563052"],"award-info":[{"award-number":["61862061","62061045","61563052"]}]}],"content-domain":{"domain":[],"crossmark-restriction":false},"short-container-title":["Sensors"],"abstract":"<jats:p>Scene text detection task aims to precisely localize text in natural environments. At present, the application scenarios of text detection topics have gradually shifted from plain document text to more complex natural scenarios. Objects with similar texture and text morphology in the complex background noise of natural scene images are prone to false recall and difficult to detect multi-scale texts, a multi-directional scene Uyghur text detection model based on fine-grained feature representation and spatial feature fusion is proposed, and feature extraction and feature fusion are improved to enhance the network\u2019s ability to represent multi-scale features. In this method, the multiple groups of 3 \u00d7 3 convolutional feature groups that are connected like the hierarchical residual to build a residual network for feature extraction, which captures the feature details and increases the receptive field of the network to adapt to multi-scale text and long glued dimensional font detection and suppress false positives of text-like objects. Secondly, an adaptive multi-level feature map fusion strategy is adopted to overcome the inconsistency of information in multi-scale feature map fusion. The proposed model achieves 93.94% and 84.92% F-measure on the self-built Uyghur dataset and the ICDAR2015 dataset, respectively, which improves the accuracy of Uyghur text detection and suppresses false positives.<\/jats:p>","DOI":"10.3390\/s22124372","type":"journal-article","created":{"date-parts":[[2022,6,13]],"date-time":"2022-06-13T02:01:44Z","timestamp":1655085704000},"page":"4372","update-policy":"https:\/\/doi.org\/10.3390\/mdpi_crossmark_policy","source":"Crossref","is-referenced-by-count":10,"title":["Scene Uyghur Text Detection Based on Fine-Grained Feature Representation"],"prefix":"10.3390","volume":"22","author":[{"given":"Yiwen","family":"Wang","sequence":"first","affiliation":[{"name":"School of Information Science and Engineering, Xinjiang University, Urumqi 830046, China"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Hornisa","family":"Mamat","sequence":"additional","affiliation":[{"name":"School of Information Science and Engineering, Xinjiang University, Urumqi 830046, China"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Xuebin","family":"Xu","sequence":"additional","affiliation":[{"name":"School of Information Science and Engineering, Xinjiang University, Urumqi 830046, China"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"ORCID":"https:\/\/orcid.org\/0000-0002-5464-0594","authenticated-orcid":false,"given":"Alimjan","family":"Aysa","sequence":"additional","affiliation":[{"name":"School of Information Science and Engineering, Xinjiang University, Urumqi 830046, China"},{"name":"Xinjiang Key Laboratory of Multilingual Information Technology, Xinjiang University, Urumqi 830046, China"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Kurban","family":"Ubul","sequence":"additional","affiliation":[{"name":"School of Information Science and Engineering, Xinjiang University, Urumqi 830046, China"},{"name":"Xinjiang Key Laboratory of Multilingual Information Technology, Xinjiang University, Urumqi 830046, China"}],"role":[{"role":"author","vocabulary":"crossref"}]}],"member":"1968","published-online":{"date-parts":[[2022,6,9]]},"reference":[{"key":"ref_1","doi-asserted-by":"crossref","first-page":"143","DOI":"10.1007\/s10032-019-00320-5","article-title":"Scene Text Detection and Recognition with Advances in Deep Learning: A Survey","volume":"22","author":"Liu","year":"2019","journal-title":"Int. J. Doc. Anal. Recognit."},{"key":"ref_2","doi-asserted-by":"crossref","unstructured":"Rani, N., Pruthvi, T., Rao, A., and Bipin, N. (2021, January 11\u201312). Automated Text Line Segmentation and Table Detection for Pre-Printed Document Image Analysis Systems. Proceedings of the 3rd International Conference on Signal Processing and Communication, Singapore.","DOI":"10.1109\/ICSPC51351.2021.9451779"},{"key":"ref_3","doi-asserted-by":"crossref","first-page":"267","DOI":"10.1007\/s10032-020-00358-w","article-title":"Detect GAN: GAN-Based Text Detector for Camera-Captured Document Images","volume":"23","author":"Zhao","year":"2020","journal-title":"Int. J. Doc. Anal. Recognit."},{"key":"ref_4","doi-asserted-by":"crossref","unstructured":"Bulatov, K., Fedotova, N., and Arlazarov, V. (2020, January 2\u20136). An Approach to Road Scene Text Recognition with Per-Frame Accumulation and Dynamic Stopping Decision. Proceedings of the Thirteenth International Conference on Machine Vision, Rome, Italy.","DOI":"10.1117\/12.2586912"},{"key":"ref_5","doi-asserted-by":"crossref","first-page":"69","DOI":"10.1504\/IJSNET.2021.113626","article-title":"Detection and Recognition of Text Traffic Signs above the Road","volume":"35","author":"Sun","year":"2021","journal-title":"Int. J. Sens. Netw."},{"key":"ref_6","unstructured":"Simonyan, K., and Zisserman, A. (2014). Very Deep Convolutional Networks for Large-Scale Image Recognition. arXiv."},{"key":"ref_7","doi-asserted-by":"crossref","unstructured":"He, K., Zhang, X., Ren, S., and Sun, J. (2016, January 26\u201330). Deep Residual Learning for Image Recognition. Proceedings of the IEEE\/CVF Conference on Computer Vision and Pattern Recognition, Las Vegas, NV, USA.","DOI":"10.1109\/CVPR.2016.90"},{"key":"ref_8","doi-asserted-by":"crossref","first-page":"161","DOI":"10.1007\/s11263-020-01369-0","article-title":"Scene Text Detection and Recognition: The Deep Learning Era","volume":"129","author":"Long","year":"2021","journal-title":"Int. J. Comput. Vis."},{"key":"ref_9","doi-asserted-by":"crossref","unstructured":"Zhong, Z., Sun, L., and Huo, Q. (2017, January 9\u201315). Improved Localization Accuracy by LocNet for Faster R-CNN Based text Detection. Proceedings of the 14th IAPR International Conference on Document Analysis and Recognition (ICDAR), Kyoto, Japan.","DOI":"10.1109\/ICDAR.2017.155"},{"key":"ref_10","doi-asserted-by":"crossref","unstructured":"Wang, Y., Xie, H., Zha, Z., Xing, M., Fu, Z., and Zhang, Y. (2020, January 14\u201319). Contournet: Taking a Further Step Toward Accurate Arbitrary-Shaped Scene Text Detection. Proceedings of the IEEE\/CVF Conference on Computer Vision and Pattern Recognition, Seattle, WA, USA.","DOI":"10.1109\/CVPR42600.2020.01177"},{"key":"ref_11","doi-asserted-by":"crossref","first-page":"1137","DOI":"10.1109\/TPAMI.2016.2577031","article-title":"Faster R-CNN: Towards Real-Time Object Detection with Region Proposal Networks","volume":"39","author":"Ren","year":"2017","journal-title":"IEEE Trans. Pattern Anal. Mach. Intell."},{"key":"ref_12","doi-asserted-by":"crossref","first-page":"3111","DOI":"10.1109\/TMM.2018.2818020","article-title":"Arbitrary-Oriented Scene Text Detection Via Rotation Proposals","volume":"20","author":"Ma","year":"2018","journal-title":"IEEE Trans. Multimed."},{"key":"ref_13","doi-asserted-by":"crossref","first-page":"106954","DOI":"10.1016\/j.patcog.2019.06.020","article-title":"Seglink++: Detecting Dense and Arbitrary-Shaped Scene Text by Instance-Aware Component Grouping","volume":"96","author":"Tang","year":"2019","journal-title":"Pattern Recognit."},{"key":"ref_14","doi-asserted-by":"crossref","unstructured":"Long, S., Ruan, J., Zhang, W., He, X., Wu, W., and Yao, C. (2018, January 8\u201314). TextSnake: A Flexible Representation for Detecting Text of Arbitrary Shapes. Proceedings of the 15th European Conference on Computer Vision, Munich, Germany.","DOI":"10.1007\/978-3-030-01216-8_2"},{"key":"ref_15","doi-asserted-by":"crossref","unstructured":"Zhou, X., Yao, C., Wen, H., Wang, Y., Zhou, S., He, W., and Liang, J. (2017, January 21\u201326). EAST: An Efficient and Accurate Scene Text Detector. Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition, Honolulu, HI, USA.","DOI":"10.1109\/CVPR.2017.283"},{"key":"ref_16","doi-asserted-by":"crossref","unstructured":"He, M., Liao, M., Yang, Z., Zhong, H., Tang, J., Cheng, W., and Bai, X. (2021, January 19\u201325). MOST: A Multi-Oriented Scene Text Detector with Localization Refinement. Proceedings of the IEEE\/CVF Conference on Computer Vision and Pattern Recognition, Web Meeting.","DOI":"10.1109\/CVPR46437.2021.00870"},{"key":"ref_17","doi-asserted-by":"crossref","unstructured":"Xiao, L., Zhou, P., Xu, K., and Zhao, X. (2021). Multi-Directional Scene Text Detection Based on Improved YOLOv3. Sensors, 21.","DOI":"10.3390\/s21144870"},{"key":"ref_18","doi-asserted-by":"crossref","unstructured":"Zhu, Y., Chen, J., Liang, L., Kuang, Z., Jin, L., and Zhang, W. (2021, January 19\u201325). Fourier Contour Embedding for Arbitrary-Shaped Text Detection. Proceedings of the IEEE\/CVF Conference on Computer Vision and Pattern Recognition, Web Meeting.","DOI":"10.1109\/CVPR46437.2021.00314"},{"key":"ref_19","doi-asserted-by":"crossref","unstructured":"Zhao, F., Shao, S., Zhang, L., and Wen, Z. (2021). A Straightforward and Efficient Instance-Aware Curved Text Detector. Sensors, 21.","DOI":"10.3390\/s21061945"},{"key":"ref_20","unstructured":"Wang, W., Xie, E., Song, X., Zang, Y., Wang, W., Lu, T., and Shen, C. (November, January 27). Efficient and Accurate Arbitrary-Shaped Text Detection with Pixel Aggregation Network. Proceedings of the IEEE\/CVF International Conference on Computer Vision, Seoul, Korea."},{"key":"ref_21","doi-asserted-by":"crossref","unstructured":"Wang, W., Xie, E., Li, X., Hou, W., Lu, T., Yu, G., and Shao, S. (2019, January 15\u201321). Shape Robust Text Detection with Progressive scale Expansion Network. Proceedings of the IEEE\/CVF Conference on Computer Vision and Pattern Recognition, Long Beach, CA, USA.","DOI":"10.1109\/CVPR.2019.00956"},{"key":"ref_22","doi-asserted-by":"crossref","unstructured":"Deng, D., Liu, H., Li, X., and Cai, D. (2018, January 2\u20137). PixelLink: Detecting Scene Text Via Instance Segmentation. Proceedings of the AAAI Conference on Artificial Intelligence, New Orleans, LA, USA.","DOI":"10.1609\/aaai.v32i1.12269"},{"key":"ref_23","doi-asserted-by":"crossref","unstructured":"Li, S., and Cao, W. (2021). SEMPANet: A Modified Path Aggregation Network with Squeeze-Excitation for Scene Text Detection. Sensors, 21.","DOI":"10.3390\/s21082657"},{"key":"ref_24","doi-asserted-by":"crossref","first-page":"5566","DOI":"10.1109\/TIP.2019.2900589","article-title":"Textfield: Learning a Deep Direction Field for Irregular Scene Text Detection","volume":"18","author":"Xu","year":"2019","journal-title":"IEEE Trans. Image Process."},{"key":"ref_25","doi-asserted-by":"crossref","unstructured":"Qin, X., Zhou, Y., Guo, Y., Wu, D., Tian, Z., Jiang, N., Wang, H., and Wang, W. (2021, January 20\u201324). Mask Is All You Need: Rethinking Mask R-CNN for Dense and Arbitrary-Shaped Scene Text Detection. Proceedings of the 29th ACM International Conference on Multimedia, Lisbon, Portugal.","DOI":"10.1145\/3474085.3475178"},{"key":"ref_26","doi-asserted-by":"crossref","unstructured":"Zhang, S., Zhu, X., Yang, C., Wang, H., and Yin, X. (2021, January 11\u201317). Adaptive Boundary Proposal Network for Arbitrary Shape Text Detection. Proceedings of the IEEE\/CVF International Conference on Computer Vision, Web Meeting.","DOI":"10.1109\/ICCV48922.2021.00134"},{"key":"ref_27","doi-asserted-by":"crossref","unstructured":"Liao, M., Wan, Z., Yao, C., Chen, K., and Bai, X. (2020, January 7\u201312). Real-Time Scene Text Detection with Differentiable Binarization. Proceedings of the AAAI Conference on Artificial Intelligence, New York, NY, USA.","DOI":"10.1609\/aaai.v34i07.6812"},{"key":"ref_28","doi-asserted-by":"crossref","first-page":"3389","DOI":"10.1109\/TMM.2018.2838320","article-title":"A fast Uyghur Text Detector for Complex Background Images","volume":"20","author":"Yan","year":"2018","journal-title":"IEEE Trans. Multimed."},{"key":"ref_29","doi-asserted-by":"crossref","first-page":"15083","DOI":"10.1007\/s11042-017-4538-8","article-title":"Detecting Uyghur Text in Complex Background Images with Convolutional Neural Network","volume":"76","author":"Fang","year":"2017","journal-title":"Multimed. Tools Appl."},{"key":"ref_30","doi-asserted-by":"crossref","first-page":"652","DOI":"10.1109\/TPAMI.2019.2938758","article-title":"Res2net: A New Multi-Scale Backbone Architecture","volume":"43","author":"Gao","year":"2019","journal-title":"IEEE Trans. Pattern Anal. Mach. Intell."},{"key":"ref_31","unstructured":"Liu, S., Huang, D., and Wang, Y. (2019). Learning Spatial Fusion for Single-Shot Object Detection. arXiv."},{"key":"ref_32","doi-asserted-by":"crossref","unstructured":"Lin, T., Piotr, D., Girshick, R., He, K., Hariharan, B., and Belongie, S. (2017, January 21\u201326). Feature Pyramid Networks for Object Detection. Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition, Honolulu, HI, USA.","DOI":"10.1109\/CVPR.2017.106"},{"key":"ref_33","doi-asserted-by":"crossref","unstructured":"Karatzas, D., Gomez-Bigorda, L., Nicolaou, A., Ghosh, S., Bagdanov, A., Iwamura, M., and Valveny, E. (2015, January 23\u201326). ICDAR 2015 Competition on Robust Reading. Proceedings of the 13th International Conference on Document Analysis and Recognition (ICDAR), Nancy, France.","DOI":"10.1109\/ICDAR.2015.7333942"}],"container-title":["Sensors"],"original-title":[],"language":"en","link":[{"URL":"https:\/\/www.mdpi.com\/1424-8220\/22\/12\/4372\/pdf","content-type":"unspecified","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2025,10,10]],"date-time":"2025-10-10T23:26:54Z","timestamp":1760138814000},"score":1,"resource":{"primary":{"URL":"https:\/\/www.mdpi.com\/1424-8220\/22\/12\/4372"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2022,6,9]]},"references-count":33,"journal-issue":{"issue":"12","published-online":{"date-parts":[[2022,6]]}},"alternative-id":["s22124372"],"URL":"https:\/\/doi.org\/10.3390\/s22124372","relation":{},"ISSN":["1424-8220"],"issn-type":[{"type":"electronic","value":"1424-8220"}],"subject":[],"published":{"date-parts":[[2022,6,9]]}}}