{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2026,6,4]],"date-time":"2026-06-04T09:08:49Z","timestamp":1780564129526,"version":"3.54.1"},"reference-count":16,"publisher":"Oxford University Press (OUP)","issue":"2","license":[{"start":{"date-parts":[[2019,5,23]],"date-time":"2019-05-23T00:00:00Z","timestamp":1558569600000},"content-version":"vor","delay-in-days":0,"URL":"https:\/\/academic.oup.com\/journals\/pages\/open_access\/funder_policies\/chorus\/standard_publication_model"}],"content-domain":{"domain":[],"crossmark-restriction":false},"short-container-title":[],"published-print":{"date-parts":[[2020,6,1]]},"abstract":"<jats:title>Abstract<\/jats:title>\n               <jats:p>Many classical texts are available in multiple versions that almost always differ from each other due to transcription error and editorial discretion. One of the central challenges in the study of such texts is the preparation of a \u2018synoptic\u2019 text: an aligned presentation of the various versions in which corresponding words or phrases, even if not identical, are mapped to each other. Multiple text alignment of this sort must take into account orthographic and conceptual relationships between words. In this article, we define this text alignment problem as an optimization problem by providing a formal measure of alignment quality. Unlike previous measures, our measure uses word embeddings to take into account conceptual similarity between aligned words. We propose an efficient and scalable alignment method in accordance with the proposed criteria. This method splits the texts to be aligned into smaller subtexts, thus improving both efficiency and accuracy. Empirical comparisons on sample data indicate our method is significantly faster than existing methods, often rendering intractable problems tractable, and that the alignment obtained by our method is considerably better than that obtained by other methods.<\/jats:p>","DOI":"10.1093\/llc\/fqz029","type":"journal-article","created":{"date-parts":[[2019,4,3]],"date-time":"2019-04-03T03:28:56Z","timestamp":1554262136000},"page":"254-264","source":"Crossref","is-referenced-by-count":4,"title":["FAST: Fast and Accurate Synoptic Texts"],"prefix":"10.1093","volume":"35","author":[{"given":"Oran","family":"Brill","sequence":"first","affiliation":[{"name":"Department of Computer Science, Bar-Ilan University, Israel"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Moshe","family":"Koppel","sequence":"first","affiliation":[{"name":"Department of Computer Science, Bar-Ilan University, Israel"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Avi","family":"Shmidman","sequence":"additional","affiliation":[{"name":"Department of Hebrew Literature, Bar-Ilan University, Israel"}],"role":[{"vocabulary":"crossref","role":"author"}]}],"member":"286","published-online":{"date-parts":[[2019,5,23]]},"reference":[{"key":"2020052917400737900_fqz029-B1","article-title":"Practical block sequence alignment with moves","volume":"3","author":"Bourdaillet","year":"2007","journal-title":"LATA 2007 - International Conference on Language and Automata Theory and Applications"},{"key":"2020052917400737900_fqz029-B2","first-page":"17","volume-title":"Supporting Digital Humanities 2011","author":"Dekker","year":"2011"},{"issue":"3","key":"2020052917400737900_fqz029-B3","doi-asserted-by":"crossref","first-page":"452","DOI":"10.1093\/llc\/fqu007","article-title":"Computer-supported collation of modern manuscripts: CollateX and the Beckett Digital Manuscript Project","volume":"30","author":"Dekker","year":"2014","journal-title":"Digital Scholarship in the Humanities"},{"key":"2020052917400737900_fqz029-B4","author":"Eger","year":"2015"},{"key":"2020052917400737900_fqz029-B5","volume-title":"Talmud Arukh: BT Bava Mezi'a VI","author":"Friedman","year":"1996"},{"issue":"14","key":"2020052917400737900_fqz029-B6","doi-asserted-by":"crossref","first-page":"3059","DOI":"10.1093\/nar\/gkf436","article-title":"MAFFT: a novel method for rapid multiple sequence alignment based on fast Fourier transform","volume":"30","author":"Katoh","year":"2002","journal-title":"Nucleic Acids Research"},{"issue":"8","key":"2020052917400737900_fqz029-B7","first-page":"707","article-title":"Binary codes capable of correcting deletions, insertions, and reversals","volume":"10","author":"Levenshtein","year":"1966","journal-title":"Soviet Physics Doklady"},{"key":"2020052917400737900_fqz029-B8","first-page":"302","article-title":"Dependency-based word embeddings","author":"Levy","year":"2014","journal-title":"Proceedings of ACL"},{"key":"2020052917400737900_fqz029-B9","article-title":"Efficient estimation of word representations in vector space","author":"Mikolov","year":"2013","journal-title":"ICLR Workshop"},{"issue":"3","key":"2020052917400737900_fqz029-B10","doi-asserted-by":"crossref","first-page":"443","DOI":"10.1016\/0022-2836(70)90057-4","article-title":"A general method applicable to the search for similarities in the amino acid sequence of two proteins","volume":"48","author":"Needleman","year":"1970","journal-title":"Journal of Molecular Biology"},{"issue":"1","key":"2020052917400737900_fqz029-B11","doi-asserted-by":"crossref","first-page":"19","DOI":"10.1162\/089120103321337421","article-title":"A systematic comparison of various statistical alignment models","volume":"29","author":"Och","year":"2003","journal-title":"Computational Linguistics"},{"key":"2020052917400737900_fqz029-B12","volume-title":"Geniza-Fragmente zur Hekhalot-Literatur.","author":"Sch\u00e4fer","year":"1984"},{"issue":"6","key":"2020052917400737900_fqz029-B13","doi-asserted-by":"crossref","first-page":"497","DOI":"10.1016\/j.ijhcs.2009.02.001","article-title":"A data structure for representing multi-version texts online","volume":"67","author":"Schmidt","year":"2009","journal-title":"International Journal of Human-Computer Studies"},{"key":"2020052917400737900_fqz029-B14","volume-title":"The Apocryphon of John: Synopsis of Nag Hammadi Codices II, 1; III, 1; and IV, 1 with BG 8502,2. Leiden","author":"Waldstein","year":"1995"},{"key":"2020052917400737900_fqz029-B15","doi-asserted-by":"crossref","DOI":"10.1515\/9783110897029","volume-title":"The Book of Tobit: Texts from the Principal Ancient and Medieval Traditions","author":"Weeks","year":"2004"},{"key":"2020052917400737900_fqz029-B16","first-page":"448","volume-title":"Proceedings of the Eighth International Joint Conference on Natural Language Processing (Volume 2: Short Papers)","author":"Xia","year":"2017"}],"container-title":["Digital Scholarship in the Humanities"],"original-title":[],"language":"en","link":[{"URL":"http:\/\/academic.oup.com\/dsh\/article-pdf\/35\/2\/254\/33324052\/fqz029.pdf","content-type":"application\/pdf","content-version":"vor","intended-application":"syndication"},{"URL":"http:\/\/academic.oup.com\/dsh\/article-pdf\/35\/2\/254\/33324052\/fqz029.pdf","content-type":"unspecified","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2020,5,29]],"date-time":"2020-05-29T22:48:21Z","timestamp":1590792501000},"score":1,"resource":{"primary":{"URL":"https:\/\/academic.oup.com\/dsh\/article\/35\/2\/254\/5497844"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2019,5,23]]},"references-count":16,"journal-issue":{"issue":"2","published-online":{"date-parts":[[2019,5,23]]},"published-print":{"date-parts":[[2020,6,1]]}},"URL":"https:\/\/doi.org\/10.1093\/llc\/fqz029","relation":{},"ISSN":["2055-7671","2055-768X"],"issn-type":[{"value":"2055-7671","type":"print"},{"value":"2055-768X","type":"electronic"}],"subject":[],"published-other":{"date-parts":[[2020,6]]},"published":{"date-parts":[[2019,5,23]]}}}