{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2025,9,11]],"date-time":"2025-09-11T19:23:26Z","timestamp":1757618606089,"version":"3.44.0"},"reference-count":35,"publisher":"Springer Science and Business Media LLC","issue":"25","license":[{"start":{"date-parts":[[2025,6,30]],"date-time":"2025-06-30T00:00:00Z","timestamp":1751241600000},"content-version":"tdm","delay-in-days":0,"URL":"https:\/\/creativecommons.org\/licenses\/by\/4.0"},{"start":{"date-parts":[[2025,6,30]],"date-time":"2025-06-30T00:00:00Z","timestamp":1751241600000},"content-version":"vor","delay-in-days":0,"URL":"https:\/\/creativecommons.org\/licenses\/by\/4.0"}],"funder":[{"name":"Department of Biotechnology, India","award":["BT\/PR16356\/BID\/7\/596\/2016"],"award-info":[{"award-number":["BT\/PR16356\/BID\/7\/596\/2016"]}]},{"DOI":"10.13039\/501100001843","name":"Science and Engineering Research Board, India","doi-asserted-by":"crossref","award":["SUR\/2022\/002903"],"award-info":[{"award-number":["SUR\/2022\/002903"]}],"id":[{"id":"10.13039\/501100001843","id-type":"DOI","asserted-by":"crossref"}]},{"DOI":"10.13039\/100008047","name":"Carnegie Mellon University","doi-asserted-by":"crossref","id":[{"id":"10.13039\/100008047","id-type":"DOI","asserted-by":"crossref"}]}],"content-domain":{"domain":["link.springer.com"],"crossmark-restriction":false},"short-container-title":["Neural Comput &amp; Applic"],"published-print":{"date-parts":[[2025,9]]},"abstract":"<jats:title>Abstract<\/jats:title>\n          <jats:p>Scene text image super-resolution (STISR), often considered a preliminary step for scene text recognition, refers to the task of enhancing the resolution of text embedded in natural scene images and plays a vital role in various applications. Most of the existing STISR methods either leverage deep convolutional neural networks by regarding text images as natural scene images or use a text recognizer\u2019s feedback as guidance to the STISR process. However, since the text recognition is initially done on low-resolution images, it is mostly inaccurate, more so as the length of the words increases, thus degrading the super-resolution process. In this paper, we introduce DEPP which utilizes dictionary embedding (DE) based probabilistic priors calculated from a large English text corpus consisting of both alphabets and digits. The initial state and the bigram probabilities obtained are fused with the probability obtained from the recognizer, before passing it onto a single image super-resolution (SISR) block. By integrating DE as a prior and implementing a modified perceptual loss, the method effectively captures the contextual information of text, enabling more accurate super-resolution and visually pleasing results. Experimental results on the benchmark TextZoom dataset demonstrate that our DEPP framework achieves superior performance compared to most existing approaches, particularly for medium and long-length words, as measured by text recognition accuracy. Since DEPP uses the text recognition attributes to rectify or guide the super-resolution process, it makes our method more domain-inspired and task-aware, compared to usual black box deep learners.<\/jats:p>","DOI":"10.1007\/s00521-025-11441-w","type":"journal-article","created":{"date-parts":[[2025,6,30]],"date-time":"2025-06-30T07:25:07Z","timestamp":1751268307000},"page":"21259-21273","update-policy":"https:\/\/doi.org\/10.1007\/springer_crossmark_policy","source":"Crossref","is-referenced-by-count":0,"title":["DEPP: dictionary embedded probabilistic priors for scene text image super-resolution"],"prefix":"10.1007","volume":"37","author":[{"ORCID":"https:\/\/orcid.org\/0000-0002-9567-1346","authenticated-orcid":false,"given":"Avigyan","family":"Bhattacharya","sequence":"first","affiliation":[],"role":[{"role":"author","vocabulary":"crossref"}]},{"ORCID":"https:\/\/orcid.org\/0000-0003-1780-0461","authenticated-orcid":false,"given":"Subhadip","family":"Basu","sequence":"additional","affiliation":[],"role":[{"role":"author","vocabulary":"crossref"}]},{"ORCID":"https:\/\/orcid.org\/0000-0002-5597-908X","authenticated-orcid":false,"given":"Tapabrata","family":"Chakraborti","sequence":"additional","affiliation":[],"role":[{"role":"author","vocabulary":"crossref"}]}],"member":"297","published-online":{"date-parts":[[2025,6,30]]},"reference":[{"issue":"5","key":"11441_CR1","doi-asserted-by":"publisher","first-page":"1063","DOI":"10.1109\/TMM.2016.2638622","volume":"19","author":"S Karaoglu","year":"2016","unstructured":"Karaoglu S, Tao R, Gevers T, Smeulders AW (2016) Words matter: Scene text for image classification and retrieval. IEEE Trans Multimed 19(5):1063\u20131076","journal-title":"IEEE Trans Multimed"},{"issue":"2","key":"11441_CR2","doi-asserted-by":"publisher","first-page":"237","DOI":"10.1016\/j.cviu.2004.02.007","volume":"96","author":"C-Y Fang","year":"2004","unstructured":"Fang C-Y, Fuh C-S, Yen P, Cherng S, Chen S-W (2004) An automatic road sign recognition system based on a computational model of human recognition processing. Comput Vis Image Understand 96(2):237\u2013268","journal-title":"Comput Vis Image Understand"},{"key":"11441_CR3","doi-asserted-by":"crossref","unstructured":"Silva SM, Jung CR (2018) License plate detection and recognition in unconstrained scenarios. In: Proceedings of ECCV, pp 580\u2013596","DOI":"10.1007\/978-3-030-01258-8_36"},{"key":"11441_CR4","doi-asserted-by":"crossref","unstructured":"He P, Huang W, Qiao Y, Loy C, Tang X (2016) Reading scene text in deep convolutional sequences. In: Proceedings of the AAAI conference on artificial intelligence, vol 30","DOI":"10.1609\/aaai.v30i1.10465"},{"key":"11441_CR5","doi-asserted-by":"crossref","unstructured":"Jaderberg M, Vedaldi A, Zisserman A (2014) Deep features for text spotting. In: Computer vision\u2013ECCV 2014: 13th European conference, Zurich, Switzerland, September 6-12, 2014, Proceedings, Part IV 13, Springer, pp 512\u2013528","DOI":"10.1007\/978-3-319-10593-2_34"},{"key":"11441_CR6","doi-asserted-by":"publisher","first-page":"1","DOI":"10.1007\/s11263-015-0823-z","volume":"116","author":"M Jaderberg","year":"2016","unstructured":"Jaderberg M, Simonyan K, Vedaldi A, Zisserman A (2016) Reading text in the wild with convolutional neural networks. Int J Comput Vis 116:1\u201320","journal-title":"Int J Comput Vis"},{"issue":"11","key":"11441_CR7","doi-asserted-by":"publisher","first-page":"2298","DOI":"10.1109\/TPAMI.2016.2646371","volume":"39","author":"B Shi","year":"2016","unstructured":"Shi B, Bai X, Yao C (2016) An end-to-end trainable neural network for image-based sequence recognition and its application to scene text recognition. IEEE Trans Pattern Anal Mach Intell 39(11):2298\u20132304","journal-title":"IEEE Trans Pattern Anal Mach Intell"},{"issue":"8","key":"11441_CR8","doi-asserted-by":"publisher","first-page":"5374","DOI":"10.1109\/TCSVT.2022.3146240","volume":"32","author":"B Li","year":"2022","unstructured":"Li B, Tang X, Qi X, Chen Y, Li C-G, Xiao R (2022) Emu: Effective multi-hot encoding net for lightweight scene text recognition with a large character set. IEEE Trans Circuits Syst Video Technol 32(8):5374\u20135385","journal-title":"IEEE Trans Circuits Syst Video Technol"},{"issue":"9","key":"11441_CR9","doi-asserted-by":"publisher","first-page":"2035","DOI":"10.1109\/TPAMI.2018.2848939","volume":"41","author":"B Shi","year":"2018","unstructured":"Shi B, Yang M, Wang X, Lyu P, Yao C, Bai X (2018) Aster: An attentional scene text recognizer with flexible rectification. IEEE Trans Pattern Anal Mach Intell 41(9):2035\u20132048","journal-title":"IEEE Trans Pattern Anal Mach Intell"},{"issue":"4","key":"11441_CR10","doi-asserted-by":"publisher","first-page":"1145","DOI":"10.1109\/TCSVT.2018.2817642","volume":"29","author":"K Raghunandan","year":"2018","unstructured":"Raghunandan K, Shivakumara P, Roy S, Kumar GH, Pal U, Lu T (2018) Multi-script-oriented text detection and recognition in video\/scene\/born digital images. IEEE Trans Circuits Syst Video Technol 29(4):1145\u20131162","journal-title":"IEEE Trans Circuits Syst Video Technol"},{"issue":"8","key":"11441_CR11","doi-asserted-by":"publisher","first-page":"3051","DOI":"10.1109\/TCSVT.2020.3037068","volume":"31","author":"J Lei","year":"2020","unstructured":"Lei J, Zhang Z, Fan X, Yang B, Li X, Chen Y, Huang Q (2020) Deep stereoscopic image super-resolution via interaction module. IEEE Trans Circuits Syst Video Technol 31(8):3051\u20133061","journal-title":"IEEE Trans Circuits Syst Video Technol"},{"issue":"8","key":"11441_CR12","doi-asserted-by":"publisher","first-page":"5137","DOI":"10.1109\/TCSVT.2022.3153390","volume":"32","author":"A Niu","year":"2022","unstructured":"Niu A, Zhu Y, Zhang C, Sun J, Wang P, Kweon IS, Zhang Y (2022) Ms2net: Multi-scale and multi-stage feature fusion for blurred image super-resolution. IEEE Trans Circuits Syst Video Technol 32(8):5137\u20135150","journal-title":"IEEE Trans Circuits Syst Video Technol"},{"issue":"2","key":"11441_CR13","doi-asserted-by":"publisher","first-page":"295","DOI":"10.1109\/TPAMI.2015.2439281","volume":"38","author":"C Dong","year":"2015","unstructured":"Dong C, Loy CC, He K, Tang X (2015) Image super-resolution using deep convolutional networks. IEEE Trans Pattern Anal Mach Intell 38(2):295\u2013307","journal-title":"IEEE Trans Pattern Anal Mach Intell"},{"issue":"4","key":"11441_CR14","doi-asserted-by":"publisher","first-page":"600","DOI":"10.1109\/TIP.2003.819861","volume":"13","author":"Z Wang","year":"2004","unstructured":"Wang Z, Bovik AC, Sheikh HR, Simoncelli EP (2004) Image quality assessment: from error visibility to structural similarity. IEEE Trans Image Process 13(4):600\u2013612","journal-title":"IEEE Trans Image Process"},{"key":"11441_CR15","doi-asserted-by":"crossref","unstructured":"Ledig C, Theis L, Husz\u00e1r F, Caballero J, Cunningham A, Acosta A, Aitken A, Tejani A, Totz J, Wang Z et al (2017) Photo-realistic single image super-resolution using a generative adversarial network. In: Proceedings of the IEEE conference on computer vision and pattern recognition, pp 4681\u20134690","DOI":"10.1109\/CVPR.2017.19"},{"key":"11441_CR16","doi-asserted-by":"crossref","unstructured":"Wang X, Yu K, Dong C, Loy CC (2018) Recovering realistic texture in image super-resolution by deep spatial feature transform. In: Proceedings of the IEEE conference on computer vision and pattern recognition, pp 606\u2013615","DOI":"10.1109\/CVPR.2018.00070"},{"key":"11441_CR17","unstructured":"Dong C, Zhu X, Deng Y, Loy CC, Qiao Y (2015) Boosting optical character recognition: a super-resolution approach. arXiv preprint arXiv:1506.02211"},{"key":"11441_CR18","first-page":"1","volume":"32","author":"Z B\u00edlkov\u00e1","year":"2020","unstructured":"B\u00edlkov\u00e1 Z, Hradi\u0161 M (2020) Perceptual license plate super-resolution with CTC loss. Electron Imag 32:1\u20135","journal-title":"Electron Imag"},{"key":"11441_CR19","unstructured":"Wang W, Xie E, Sun P, Wang W, Tian L, Shen C, Luo P (2019) Textsr: Content-aware text super-resolution guided by recognition. arXiv preprint arXiv:1909.07113"},{"key":"11441_CR20","doi-asserted-by":"crossref","unstructured":"Graves A, Fern\u00e1ndez S, Gomez F, Schmidhuber J (2006) Connectionist temporal classification: labelling unsegmented sequence data with recurrent neural networks. In: Proceedings of the 23rd international conference on machine learning, pp 369\u2013376","DOI":"10.1145\/1143844.1143891"},{"key":"11441_CR21","unstructured":"Wang W, Xie E, Liu X, Wang W, Liang D, Shen C, Bai X (2020) Scene text image super-resolution in the wild. In: ECCV 2020: 16th European conference, Glasgow, UK, August 23\u201328, 2020, proceedings, Part X 16, Springer"},{"key":"11441_CR22","doi-asserted-by":"crossref","unstructured":"Xu X, Sun D, Pan J, Zhang Y, Pfister H, Yang M-H (2017) Learning to super-resolve blurry face and text images. In: Proceedings of the IEEE international conference on computer vision, pp 251\u2013260","DOI":"10.1109\/ICCV.2017.36"},{"key":"11441_CR23","first-page":"778","volume":"6","author":"Y Quan","year":"2020","unstructured":"Quan Y, Yang J, Chen Y, Xu Y, Ji H (2020) Collaborative deep learning for super-resolving blurry text images. IEEE Trans Computat Imag 6:778\u2013790","journal-title":"IEEE Trans Computat Imag"},{"key":"11441_CR24","doi-asserted-by":"crossref","unstructured":"Zhao C, Feng S, Zhao BN, Ding Z, Wu J, Shen F, Shen HT (2021) Scene text image super-resolution via parallelly contextual attention network. In: Proceedings of the 29th ACM international conference on multimedia, pp 2908\u20132917","DOI":"10.1145\/3474085.3475469"},{"key":"11441_CR25","doi-asserted-by":"crossref","unstructured":"Chen J, Li B, Xue X (2021) Scene text telescope: Text-focused scene image super-resolution. In: Proceedings of the IEEE\/CVF conference on computer vision and pattern recognition, pp 12026\u201312035","DOI":"10.1109\/CVPR46437.2021.01185"},{"key":"11441_CR26","doi-asserted-by":"publisher","first-page":"1341","DOI":"10.1109\/TIP.2023.3237002","volume":"32","author":"J Ma","year":"2023","unstructured":"Ma J, Guo S, Zhang L (2023) Text prior guided scene text image super-resolution. IEEE Trans Image Process 32:1341\u20131353","journal-title":"IEEE Trans Image Process"},{"key":"11441_CR27","doi-asserted-by":"crossref","unstructured":"Shi W, Caballero J, Husz\u00e1r F, Totz J, Aitken AP, Bishop R, Rueckert D, Wang Z (2016) Real-time single image and video super-resolution using an efficient sub-pixel convolutional neural network. In: Proceedings of the IEEE conference on computer vision and pattern recognition, pp 1874\u20131883","DOI":"10.1109\/CVPR.2016.207"},{"key":"11441_CR28","unstructured":"Johnson J, Alahi A, Fei-Fei L (2016) Perceptual losses for real-time style transfer and super-resolution. In: ECCV 2016: 14th European conference, Amsterdam, The Netherlands, October 11-14, 2016, proceedings, Part II 14, Springer"},{"key":"11441_CR29","doi-asserted-by":"crossref","unstructured":"Karatzas D, Gomez-Bigorda L, Nicolaou A, Ghosh S, Bagdanov A, Iwamura M, Matas J, Neumann L, Chandrasekhar VR, Lu S, et al. (2015) Icdar 2015 competition on robust reading. In: 2015 13th International conference on document analysis and recognition (ICDAR), pp 1156\u20131160. IEEE","DOI":"10.1109\/ICDAR.2015.7333942"},{"key":"11441_CR30","unstructured":"Wang K, Babenko B, Belongie S (2011) End-to-end scene text recognition. In: 2011 International conference on computer vision, IEEE, pp 1457\u20131464"},{"key":"11441_CR31","doi-asserted-by":"crossref","unstructured":"Zhang Y, Tian Y, Kong Y, Zhong B, Fu Y (2018) Residual dense network for image super-resolution. In: Proceedings of the IEEE CVPR, pp 2472\u20132481","DOI":"10.1109\/CVPR.2018.00262"},{"key":"11441_CR32","doi-asserted-by":"crossref","unstructured":"Lim B, Son S, Kim H, Nah S, Mu Lee K (2017) Enhanced deep residual networks for single image super-resolution. In: Proceedings of the IEEE conference on computer vision and pattern recognition workshops, pp 136\u2013144","DOI":"10.1109\/CVPRW.2017.151"},{"key":"11441_CR33","doi-asserted-by":"crossref","unstructured":"Lai W-S, Huang J-B, Ahuja N, Yang M-H (2017) Deep laplacian pyramid networks for fast and accurate super-resolution. In: Proceedings of the IEEE conference on computer vision and pattern recognition, pp 624\u2013632","DOI":"10.1109\/CVPR.2017.618"},{"key":"11441_CR34","doi-asserted-by":"crossref","unstructured":"Qiao Z, Zhou Y, Yang D, Zhou Y, Wang W (2020) Seed: Semantics enhanced encoder-decoder framework for scene text recognition. In: Proceedings of the IEEE\/CVF conference on computer vision and pattern recognition, pp 13528\u201313537","DOI":"10.1109\/CVPR42600.2020.01354"},{"key":"11441_CR35","doi-asserted-by":"publisher","first-page":"109","DOI":"10.1016\/j.patcog.2019.01.020","volume":"90","author":"C Luo","year":"2019","unstructured":"Luo C, Jin L, Sun Z (2019) Moran: A multi-object rectified attention network for scene text recognition. Pattern Recog 90:109\u2013118","journal-title":"Pattern Recog"}],"container-title":["Neural Computing and Applications"],"original-title":[],"language":"en","link":[{"URL":"https:\/\/link.springer.com\/content\/pdf\/10.1007\/s00521-025-11441-w.pdf","content-type":"application\/pdf","content-version":"vor","intended-application":"text-mining"},{"URL":"https:\/\/link.springer.com\/article\/10.1007\/s00521-025-11441-w\/fulltext.html","content-type":"text\/html","content-version":"vor","intended-application":"text-mining"},{"URL":"https:\/\/link.springer.com\/content\/pdf\/10.1007\/s00521-025-11441-w.pdf","content-type":"application\/pdf","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2025,9,6]],"date-time":"2025-09-06T23:59:20Z","timestamp":1757203160000},"score":1,"resource":{"primary":{"URL":"https:\/\/link.springer.com\/10.1007\/s00521-025-11441-w"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2025,6,30]]},"references-count":35,"journal-issue":{"issue":"25","published-print":{"date-parts":[[2025,9]]}},"alternative-id":["11441"],"URL":"https:\/\/doi.org\/10.1007\/s00521-025-11441-w","relation":{},"ISSN":["0941-0643","1433-3058"],"issn-type":[{"type":"print","value":"0941-0643"},{"type":"electronic","value":"1433-3058"}],"subject":[],"published":{"date-parts":[[2025,6,30]]},"assertion":[{"value":"15 February 2025","order":1,"name":"received","label":"Received","group":{"name":"ArticleHistory","label":"Article History"}},{"value":"10 June 2025","order":2,"name":"accepted","label":"Accepted","group":{"name":"ArticleHistory","label":"Article History"}},{"value":"30 June 2025","order":3,"name":"first_online","label":"First Online","group":{"name":"ArticleHistory","label":"Article History"}},{"value":"17 July 2025","order":4,"name":"change_date","label":"Change Date","group":{"name":"ArticleHistory","label":"Article History"}},{"value":"Correction","order":5,"name":"change_type","label":"Change Type","group":{"name":"ArticleHistory","label":"Article History"}},{"value":"The original online version of this article was revised to update the affiliation section.","order":6,"name":"change_details","label":"Change Details","group":{"name":"ArticleHistory","label":"Article History"}},{"order":1,"name":"Ethics","group":{"name":"EthicsHeading","label":"Declarations"}},{"value":"All the authors declare no Conflict of interest.","order":2,"name":"Ethics","group":{"name":"EthicsHeading","label":"Conflict of interest"}}]}}