{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2026,3,7]],"date-time":"2026-03-07T00:09:12Z","timestamp":1772842152147,"version":"3.50.1"},"reference-count":28,"publisher":"Oxford University Press (OUP)","issue":"1","license":[{"start":{"date-parts":[[2024,2,1]],"date-time":"2024-02-01T00:00:00Z","timestamp":1706745600000},"content-version":"vor","delay-in-days":0,"URL":"https:\/\/academic.oup.com\/pages\/standard-publication-reuse-rights"}],"funder":[{"DOI":"10.13039\/501100009328","name":"Planning and Budgeting Committee","doi-asserted-by":"publisher","id":[{"id":"10.13039\/501100009328","id-type":"DOI","asserted-by":"publisher"}]},{"DOI":"10.13039\/100019473","name":"Ministry of Science & Technology","doi-asserted-by":"publisher","award":["3-16464"],"award-info":[{"award-number":["3-16464"]}],"id":[{"id":"10.13039\/100019473","id-type":"DOI","asserted-by":"publisher"}]}],"content-domain":{"domain":[],"crossmark-restriction":false},"short-container-title":[],"published-print":{"date-parts":[[2024,4,2]]},"abstract":"<jats:title>Abstract<\/jats:title>\n               <jats:p>We describe an efficient pipeline for morpho-syntactically annotating an ancient language corpus which takes advantage of bootstrapping techniques. This pipeline is designed for ancient language scholars looking to jump-start their own treebank projects, which can in turn serve further pedagogical research projects in the target language. We situate our work in the field of similar ancient language treebank projects, arguing that our approach shows that individual humanities scholars can leverage current machine-learning tools to produce their own richly annotated corpora. We illustrate this pipeline by producing a new Akkadian-language treebank based on two volumes from the online editions of the State Archives of Assyria project hosted on Oracc, as well as a spaCy language model named AkkParser trained on that treebank. Both of these are made publicly available for annotating other Akkadian corpora. In addition, we discuss linguistic issues particular to the Neo-Assyrian letter corpus and data-encoding complications of cuneiform texts in Oracc. The strategies, language models, and processing scripts we developed to handle both linguistic and data-encoding issues in this project will be of special interest to scholars seeking to develop their own cuneiform treebanks.<\/jats:p>","DOI":"10.1093\/llc\/fqae002","type":"journal-article","created":{"date-parts":[[2024,2,2]],"date-time":"2024-02-02T11:46:43Z","timestamp":1706874403000},"page":"296-307","source":"Crossref","is-referenced-by-count":1,"title":["Linguistic annotation of cuneiform texts using treebanks and deep learning"],"prefix":"10.1093","volume":"39","author":[{"given":"Matthew","family":"Ong","sequence":"first","affiliation":[{"name":"Middle Eastern Languages and Cultures, UC Berkeley , Berkeley, CA, United States"}]},{"ORCID":"https:\/\/orcid.org\/0000-0002-8359-382X","authenticated-orcid":false,"given":"Shai","family":"Gordin","sequence":"additional","affiliation":[{"name":"Digital Pasts Lab, Department of Land of Israel Studies and Archaeology, Ariel University , Ariel, Israel"},{"name":"Digital Humanities and Social Sciences Hub, Open University of Israel , Ra'anana, Israel"}]}],"member":"286","published-online":{"date-parts":[[2024,2,1]]},"reference":[{"key":"2024040210383158700_fqae002-B1","doi-asserted-by":"crossref","first-page":"79","DOI":"10.1007\/978-3-642-20227-8_5","volume-title":"Language Technology for Cultural Heritage","author":"Bamman","year":"2011"},{"key":"2024040210383158700_fqae002-B2","doi-asserted-by":"crossref","first-page":"1","DOI":"10.1163\/26670755-01010003","article-title":"\u2018Evaluating Syntactic Annotation of Ancient Languages: Lessons from the Vedic Treebank","volume":"1","author":"Biagetti","year":"2021","journal-title":"Old World: Journal of Ancient Africa and Eurasia"},{"key":"2024040210383158700_fqae002-B3","first-page":"1","author":"Dukes","year":"2010"},{"key":"2024040210383158700_fqae002-B4","doi-asserted-by":"crossref","first-page":"154","DOI":"10.1075\/bct.113","volume-title":"Diachronic Treebanks for Historical Linguistics","author":"Eckhoff","year":"2020"},{"key":"2024040210383158700_fqae002-B5","doi-asserted-by":"crossref","first-page":"29","DOI":"10.1007\/s10579-017-9388-5","article-title":"\u2018The PROIEL Treebank Family: A Standard for Early Attestations of Indo-European Languages","volume":"52","author":"Eckhoff","year":"2018","journal-title":"Language Resources and Evaluation"},{"key":"2024040210383158700_fqae002-B6","doi-asserted-by":"crossref","first-page":"026","DOI":"10.17352\/tcsit.000048","article-title":"\u2018AI Trends in Digital Humanities Research","volume":"7","author":"George","year":"2022","journal-title":"Trends in Computer Science and Information Technology"},{"key":"2024040210383158700_fqae002-B7","first-page":"1","volume-title":"Proceedings of the Sixth International Conference on Dependency Linguistics (Depling, SyntaxFest 2021)","author":"Al-Ghamdi","year":"2021"},{"key":"2024040210383158700_fqae002-B8","first-page":"1","article-title":"\u2018Seeing the Light through the Trees: How Treebanks Can Advance the Education of Classical Languages","volume":"89","author":"Hal","year":"2021","journal-title":"Les \u00c9tudes Classiques"},{"key":"2024040210383158700_fqae002-B9","first-page":"5137","author":"Hellwig","year":"2020"},{"key":"2024040210383158700_fqae002-B10","doi-asserted-by":"crossref","first-page":"295","DOI":"10.1075\/cf.8.2.06hon","article-title":"\u2018Automatic Metaphor Detection Using Constructions and Frames","volume":"8","author":"Hong","year":"2016","journal-title":"Constructions and Frames"},{"key":"2024040210383158700_fqae002-B11","first-page":"20","author":"Johnson","year":"2021"},{"key":"2024040210383158700_fqae002-B12","first-page":"133","author":"Kanerva","year":"2018"},{"key":"2024040210383158700_fqae002-B13","first-page":"59","author":"Keersmaekers","year":"2020"},{"key":"2024040210383158700_fqae002-B14","first-page":"5","author":"Klie","year":"2018"},{"key":"2024040210383158700_fqae002-B15","volume-title":"Metaphors We Live By","author":"Lakoff","year":"1980"},{"key":"2024040210383158700_fqae002-B16","first-page":"191","author":"Lee","year":"2012"},{"key":"2024040210383158700_fqae002-B17","first-page":"124","author":"Luukko","year":"2020"},{"key":"2024040210383158700_fqae002-B18","doi-asserted-by":"crossref","first-page":"1","DOI":"10.1163\/24523666-06010004","article-title":"\u2018The DiGreC Treebank: Linguistics and Literature","volume":"6","author":"Macleod","year":"2021","journal-title":"Research Data Journal for the Humanities and Social Sciences"},{"key":"2024040210383158700_fqae002-B19","author":"Mambrini","year":"2016"},{"key":"2024040210383158700_fqae002-B20","doi-asserted-by":"crossref","first-page":"299","DOI":"10.1515\/9783110599572-017","volume-title":"Digital Classical Philology. Ancient Greek and Latin in the Digital Revolution","author":"Passarotti","year":"2019"},{"key":"2024040210383158700_fqae002-B21","first-page":"101","author":"Qi","year":"2020"},{"key":"2024040210383158700_fqae002-B22","first-page":"14","author":"Sahala","year":"2022"},{"key":"2024040210383158700_fqae002-B23","first-page":"3886","author":"Sahala","year":"2020"},{"key":"2024040210383158700_fqae002-B24","doi-asserted-by":"crossref","first-page":"193","DOI":"10.1075\/cal.14","volume-title":"Frames and Constructions in Metaphoric Language, p.","author":"Sullivan","year":"2013"},{"key":"2024040210383158700_fqae002-B25","first-page":"2353","author":"Swanson","year":"2022"},{"key":"2024040210383158700_fqae002-B26","doi-asserted-by":"crossref","first-page":"5","DOI":"10.1007\/978-94-010-0201-1_1","volume-title":"Treebanks: Building and Using Parsed Corpora","author":"Taylor","year":"2003"},{"key":"2024040210383158700_fqae002-B27","first-page":"416","volume-title":"Grundriss der Akkadischen Grammatik","author":"Von Soden","year":"1995","edition":"3rd ed"},{"issue":"1","key":"2024040210383158700_fqae002-B28","doi-asserted-by":"crossref","first-page":"012011","DOI":"10.1088\/1757-899X\/806\/1\/012011","article-title":"Mapping Big Data and Artificial Intelligence in arts and humanities across time: an exploratory scientometric analysis based on the WoS database\u2019","volume":"806","author":"Xu,","journal-title":"IOP Conference Series: Materials Science and Engineering"}],"container-title":["Digital Scholarship in the Humanities"],"original-title":[],"language":"en","link":[{"URL":"https:\/\/academic.oup.com\/dsh\/article-pdf\/39\/1\/296\/57134637\/fqae002.pdf","content-type":"application\/pdf","content-version":"vor","intended-application":"syndication"},{"URL":"https:\/\/academic.oup.com\/dsh\/article-pdf\/39\/1\/296\/57134637\/fqae002.pdf","content-type":"unspecified","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2024,4,2]],"date-time":"2024-04-02T13:57:37Z","timestamp":1712066257000},"score":1,"resource":{"primary":{"URL":"https:\/\/academic.oup.com\/dsh\/article\/39\/1\/296\/7596406"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2024,2,1]]},"references-count":28,"journal-issue":{"issue":"1","published-online":{"date-parts":[[2024,2,1]]},"published-print":{"date-parts":[[2024,4,2]]}},"URL":"https:\/\/doi.org\/10.1093\/llc\/fqae002","relation":{},"ISSN":["2055-7671","2055-768X"],"issn-type":[{"value":"2055-7671","type":"print"},{"value":"2055-768X","type":"electronic"}],"subject":[],"published-other":{"date-parts":[[2024,4,1]]},"published":{"date-parts":[[2024,2,1]]}}}