{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2025,10,11]],"date-time":"2025-10-11T01:45:13Z","timestamp":1760147113230,"version":"build-2065373602"},"reference-count":39,"publisher":"MDPI AG","issue":"1","license":[{"start":{"date-parts":[[2023,1,13]],"date-time":"2023-01-13T00:00:00Z","timestamp":1673568000000},"content-version":"vor","delay-in-days":0,"URL":"https:\/\/creativecommons.org\/licenses\/by\/4.0\/"}],"funder":[{"name":"Department of Information and Electrical Engineering and Applied Mathematics of the University of Salerno (Italy)"}],"content-domain":{"domain":[],"crossmark-restriction":false},"short-container-title":["J. Imaging"],"abstract":"<jats:p>The growth of digital libraries has yielded a large number of handwritten historical documents in the form of images, often accompanied by a digital transcription of the content. The ability to track the position of the words of the digital transcription in the images can be important both for the study of the document by humanities scholars and for further automatic processing. We propose a learning-free method for automatically aligning the transcription to the document image. The method receives as input the digital image of the document and the transcription of its content and aims at linking the transcription to the corresponding images within the page at the word level. The method comprises two main original contributions: a line-level segmentation algorithm capable of detecting text lines with curved baseline, and a text-to-image alignment algorithm capable of dealing with under- and over-segmentation errors at the word level. Experiments on pages from a 17th-century Italian manuscript have demonstrated that the line segmentation method allows one to segment 92% of the text line correctly. They also demonstrated that it achieves a correct alignment accuracy greater than 68%. Moreover, the performance achieved on widely used data sets compare favourably with the state of the art.<\/jats:p>","DOI":"10.3390\/jimaging9010017","type":"journal-article","created":{"date-parts":[[2023,1,13]],"date-time":"2023-01-13T05:09:32Z","timestamp":1673586572000},"page":"17","update-policy":"https:\/\/doi.org\/10.3390\/mdpi_crossmark_policy","source":"Crossref","is-referenced-by-count":3,"title":["End-to-End Transcript Alignment of 17th Century Manuscripts: The Case of Moccia Code"],"prefix":"10.3390","volume":"9","author":[{"ORCID":"https:\/\/orcid.org\/0000-0002-8195-4118","authenticated-orcid":false,"given":"Giuseppe","family":"De Gregorio","sequence":"first","affiliation":[{"name":"Department of Information and Electrical Engineering and Applied Mathematics, University of Salerno, Via Giovanni Paolo II, 132, 84084 Fisciano, Italy"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"ORCID":"https:\/\/orcid.org\/0000-0003-4312-0814","authenticated-orcid":false,"given":"Giuliana","family":"Capriolo","sequence":"additional","affiliation":[{"name":"Department of Cultural Heritage, University of Salerno, Via Giovanni Paolo II, 132, 84084 Fisciano, Italy"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"ORCID":"https:\/\/orcid.org\/0000-0002-2019-2826","authenticated-orcid":false,"given":"Angelo","family":"Marcelli","sequence":"additional","affiliation":[{"name":"Department of Information and Electrical Engineering and Applied Mathematics, University of Salerno, Via Giovanni Paolo II, 132, 84084 Fisciano, Italy"}],"role":[{"role":"author","vocabulary":"crossref"}]}],"member":"1968","published-online":{"date-parts":[[2023,1,13]]},"reference":[{"key":"ref_1","unstructured":"(2022, August 15). DVL\u2014Digital Vatican Library. Available online: https:\/\/digi.vatlib.it."},{"key":"ref_2","unstructured":"(2022, August 15). Gallica. Available online: https:\/\/gallica.bnf.fr."},{"key":"ref_3","unstructured":"(2022, August 15). e-codices\u2014Virtual Manuscript Library of Switzerland. Available online: https:\/\/www.e-codices.unifr.ch."},{"key":"ref_4","unstructured":"(2022, August 15). Manuscripta Mediaevalia. Available online: http:\/\/www.manuscripta-mediaevalia.de\/."},{"key":"ref_5","unstructured":"Internet Culturale (2022, August 15). Cataloghi e Collezioni Digitali Delle Biblioteche Italiane. Available online: http:\/\/www.internetculturale.it."},{"key":"ref_6","doi-asserted-by":"crossref","unstructured":"Lombardi, F., and Marinai, S. (2020). Deep Learning for Historical Document Analysis and Recognition\u2014A Survey. J. Imaging, 6.","DOI":"10.3390\/jimaging6100110"},{"key":"ref_7","doi-asserted-by":"crossref","first-page":"122","DOI":"10.1016\/j.patcog.2019.05.025","article-title":"A set of benchmarks for Handwritten Text Recognition on historical documents","volume":"94","author":"Romero","year":"2019","journal-title":"Pattern Recognit."},{"key":"ref_8","doi-asserted-by":"crossref","unstructured":"Parziale, A., Capriolo, G., and Marcelli, A. (2020). One Step Is Not Enough: A Multi-Step Procedure for Building the Training Set of a Query by String Keyword Spotting System to Assist the Transcription of Historical Document. J. Imaging, 6.","DOI":"10.3390\/jimaging6100109"},{"key":"ref_9","unstructured":"Tomai, C.I., Zhang, B., and Govindaraju, V. (2002, January 6\u20138). Transcript mapping for historic handwritten document images. Proceedings of the Proceedings Eighth International Workshop on Frontiers in Handwriting Recognition, Niagara-on-the-Lake, ON, Canada."},{"key":"ref_10","unstructured":"Kornfield, E., Manmatha, R., and Allan, J. (2004, January 23\u201324). Text alignment with handwritten documents. Proceedings of the First International Workshop on Document Image Analysis for Libraries, Palo Alto, CA, USA."},{"key":"ref_11","doi-asserted-by":"crossref","unstructured":"Bunke, H., and Spitz, A.L. (2006). Aligning Transcripts to Automatically Segmented Handwritten Manuscripts. Proceedings of the Document Analysis Systems VII, Springer.","DOI":"10.1007\/11669487"},{"key":"ref_12","unstructured":"Toselli, A.H., Romero, V., and Vidal, E. (2007, January 28). Viterbi based alignment between text images and their transcripts. Proceedings of the Workshop on Language Technology for Cultural Heritage Data (LaTeCH 2007), Prague, Czech Republic."},{"key":"ref_13","doi-asserted-by":"crossref","unstructured":"Zinger, S., Nerbonne, J., and Schomaker, L. (2009, January 20\u201322). Text-image alignment for historical handwritten documents. Proceedings of the Document Recognition and Retrieval XVI, San Jose, CA, USA.","DOI":"10.1117\/12.805511"},{"key":"ref_14","doi-asserted-by":"crossref","unstructured":"Inderm\u00fchle, E., Liwicki, M., and Bunke, H. (2009, January 26\u201329). Combining alignment results for historical handwritten document analysis. Proceedings of the 2009 10th International Conference on Document Analysis and Recognition, Barcelona, Spain.","DOI":"10.1109\/ICDAR.2009.19"},{"key":"ref_15","doi-asserted-by":"crossref","unstructured":"Stamatopoulos, N., Louloudis, G., and Gatos, B. (2010, January 16\u201318). Efficient transcript mapping to ease the creation of document image segmentation ground truth with text-image alignment. Proceedings of the 2010 12th International Conference on Frontiers in Handwriting Recognition, Kolkata, India.","DOI":"10.1109\/ICFHR.2010.43"},{"key":"ref_16","doi-asserted-by":"crossref","unstructured":"Stamatopoulos, N., Gatos, B., and Louloudis, G. (2014, January 1\u20134). A Novel Transcript Mapping Technique for Handwritten Document Images. Proceedings of the 2014 14th International Conference on Frontiers in Handwriting Recognition, Crete, Greece.","DOI":"10.1109\/ICFHR.2014.15"},{"key":"ref_17","doi-asserted-by":"crossref","unstructured":"Leydier, Y., \u00c9glin, V., Br\u00e8s, S., and Stutzmann, D. (2014, January 1\u20134). Learning-Free Text-Image Alignment for Medieval Manuscripts. Proceedings of the 2014 14th International Conference on Frontiers in Handwriting Recognition, Crete, Greece.","DOI":"10.1109\/ICFHR.2014.67"},{"key":"ref_18","doi-asserted-by":"crossref","unstructured":"Romero-G\u00f3mez, V., Toselli, A.H., Bosch, V., S\u00e1nchez, J.A., and Vidal, E. (2018, January 24\u201327). Automatic alignment of handwritten images and transcripts for training handwritten text recognition systems. Proceedings of the 2018 13th IAPR International Workshop on Document Analysis Systems (DAS), Vienna, Austria.","DOI":"10.1109\/DAS.2018.41"},{"key":"ref_19","doi-asserted-by":"crossref","first-page":"109","DOI":"10.1016\/j.patrec.2020.02.016","article-title":"Text alignment in early printed books combining deep learning and dynamic programming","volume":"133","author":"Ziran","year":"2020","journal-title":"Pattern Recognit. Lett."},{"key":"ref_20","doi-asserted-by":"crossref","unstructured":"Torras, P., Souibgui, M.A., Chen, J., and Forn\u00e9s, A. (2021, January 5\u201310). A Transcription Is All You Need: Learning to Align Through Attention. Proceedings of the International Conference on Document Analysis and Recognition, Lausanne, Switzerland.","DOI":"10.1007\/978-3-030-86198-8_11"},{"key":"ref_21","unstructured":"Capriolo, G. (2017). Paternas Literas Confirmamus: Il Libro dei Privilegi e Delle Facolt\u00e0 del Mastro Portolano di Terra di Lavoro (secc. XV-XVII), FedOA-Federico II University Press."},{"key":"ref_22","doi-asserted-by":"crossref","first-page":"225","DOI":"10.1016\/S0031-3203(99)00055-2","article-title":"Adaptive document image binarization","volume":"33","author":"Sauvola","year":"2000","journal-title":"Pattern Recognit."},{"key":"ref_23","doi-asserted-by":"crossref","first-page":"647","DOI":"10.1147\/rd.266.0647","article-title":"Document Analysis System","volume":"26","author":"Wong","year":"1982","journal-title":"IBM J. Res. Dev."},{"key":"ref_24","doi-asserted-by":"crossref","first-page":"38","DOI":"10.1109\/34.824820","article-title":"Twenty years of document image analysis in PAMI","volume":"22","author":"Nagy","year":"2000","journal-title":"IEEE Trans. Patt. Anal. Mach. Intell."},{"key":"ref_25","doi-asserted-by":"crossref","unstructured":"Namboodiri, A., and Jain, A.K. (2007). Document structure and layout analysis. Proceedings of the Digital Document Processing, Springer.","DOI":"10.1007\/978-1-84628-726-8_2"},{"key":"ref_26","doi-asserted-by":"crossref","unstructured":"Kise, K. (2014). Page segmentation techniques in document analysis. Proceedings of the Handbook of Document Image Processing and Recognition, Springer.","DOI":"10.1007\/978-0-85729-859-1_5"},{"key":"ref_27","doi-asserted-by":"crossref","first-page":"1","DOI":"10.1016\/j.patcog.2016.10.023","article-title":"A comprehensive survey of mostly textual document segmentation algorithms since 2008","volume":"64","author":"Eskenazi","year":"2017","journal-title":"Pattern Recognit."},{"key":"ref_28","doi-asserted-by":"crossref","unstructured":"Antonacopoulus, A., Clausner, C., Papadopoulos, C., and Pletschacher, S. (2011, January 26\u201329). ICDAR2009 Page segmentation competition. Proceedings of the 2009 International Conference on Document Analysis and Recognition (ICDAR), Barcelona, Spain.","DOI":"10.1109\/ICDAR.2009.275"},{"key":"ref_29","doi-asserted-by":"crossref","unstructured":"Murdock, M., Reid, S., Hamilton, B., and Reese, J. (2015, January 23\u201326). ICDAR 2015 Competition on text line detection in historical documents. Proceedings of the 2015 International Conference on Document Analysis and Recognition (ICDAR), Tunis, Tunisia.","DOI":"10.1109\/ICDAR.2015.7333945"},{"key":"ref_30","doi-asserted-by":"crossref","unstructured":"Diem, M., Kleber, F., Fiel, S., Gr\u00fcning, T., and Gatos, B. (2017, January 9\u201315). cBAD: ICDAR2017 competition on baseline detection. Proceedings of the 2017 14th IAPR International Conference on Document Analysis and Recognition (ICDAR), Kyoto, Japan.","DOI":"10.1109\/ICDAR.2017.222"},{"key":"ref_31","doi-asserted-by":"crossref","unstructured":"Zhang, R., Zhou, Y., Jiang, Q., Song, Q., Li, N., Zhou, K., Wang, L., Wang, D., Liao, M., and Yang, M. (2019, January 20\u201325). Icdar 2019 robust reading challenge on reading chinese text on signboard. Proceedings of the 2019 International Conference on Document Analysis and Recognition (ICDAR), Sydney, Australia.","DOI":"10.1109\/ICDAR.2019.00253"},{"key":"ref_32","doi-asserted-by":"crossref","unstructured":"Surinta, O., Holtkamp, M., Karabaa, F., Van Oosten, J.P., Schomaker, L., and Wiering, M. (2014, January 1\u20134). A path planning for line segmentation of handwritten documents. Proceedings of the 2014 14th International Conference on Frontiers in Handwriting Recognition, Crete, Greece.","DOI":"10.1109\/ICFHR.2014.37"},{"key":"ref_33","doi-asserted-by":"crossref","unstructured":"De Gregorio, G., Citro, I., and Marcelli, A. (2022, January 7\u20139). Transcript Alignment for Historical Handwritten Documents: The MiM Algorithm. Proceedings of the 20th International Graphonomics Society Conference, Las Palmas de Gran Canaria, Spain. in press.","DOI":"10.1007\/978-3-031-19745-1_4"},{"key":"ref_34","doi-asserted-by":"crossref","unstructured":"Alberti, M., V\u00f6gtlin, L., Pondenkandath, V., Seuret, M., Ingold, R., and Liwicki, M. (2019, January 20\u201325). Labeling, cutting, grouping: An efficient text line segmentation method for medieval manuscripts. Proceedings of the 2019 International Conference on Document Analysis and Recognition (ICDAR), Sydney, Australia.","DOI":"10.1109\/ICDAR.2019.00194"},{"key":"ref_35","doi-asserted-by":"crossref","unstructured":"Monnier, T., and Aubry, M. (2020, January 7\u201310). docExtractor: An off-the-shelf historical document element extraction. Proceedings of the 2020 17th International Conference on Frontiers in Handwriting Recognition (ICFHR), Dortmund, Germany.","DOI":"10.1109\/ICFHR2020.2020.00027"},{"key":"ref_36","unstructured":"Oliveira, A., Seguin, B., and Kaplan, F. (2018, January 5\u20138). dhSegment: A Generic Deep-Learning Approach for Document Segmentation. Proceedings of the 2018 16th International Conference on Frontiers in Handwriting Recognition (ICFHR), Niagara Falls, NY, USA."},{"key":"ref_37","unstructured":"(2022, August 05). Transcribe Bentham. Available online: https:\/\/www.ucl.ac.uk\/bentham-project\/research-tools."},{"key":"ref_38","doi-asserted-by":"crossref","first-page":"139","DOI":"10.1007\/s10032-006-0027-8","article-title":"Word spotting for historical documents","volume":"9","author":"Rath","year":"2007","journal-title":"Int. J. Doc. Anal. Recognit. (IJDAR)"},{"key":"ref_39","doi-asserted-by":"crossref","unstructured":"Fischer, A., Frinken, V., Forn\u00e9s, A., and Bunke, H. (2011, January 16\u201317). Transcription alignment of Latin manuscripts using hidden Markov models. Proceedings of the 2011 Workshop on Historical Document Imaging and Processing, Beijing, China.","DOI":"10.1145\/2037342.2037348"}],"container-title":["Journal of Imaging"],"original-title":[],"language":"en","link":[{"URL":"https:\/\/www.mdpi.com\/2313-433X\/9\/1\/17\/pdf","content-type":"unspecified","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2025,10,10]],"date-time":"2025-10-10T18:05:15Z","timestamp":1760119515000},"score":1,"resource":{"primary":{"URL":"https:\/\/www.mdpi.com\/2313-433X\/9\/1\/17"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2023,1,13]]},"references-count":39,"journal-issue":{"issue":"1","published-online":{"date-parts":[[2023,1]]}},"alternative-id":["jimaging9010017"],"URL":"https:\/\/doi.org\/10.3390\/jimaging9010017","relation":{},"ISSN":["2313-433X"],"issn-type":[{"type":"electronic","value":"2313-433X"}],"subject":[],"published":{"date-parts":[[2023,1,13]]}}}