{"title":"Toward the morpho-syntactic annotation of an Old English corpus with universal dependencies","authors":"Javier Martín Arista","doi":"10.4995/rlyla.2022.16787","DOIUrl":null,"url":null,"abstract":"The aim of this article is to take the first steps toward the compilation of a treebank of Old English compatible with the framework of Universal Dependencies (UD). Such a treebank will comprise morphological and syntactic annotation of Old English texts adequate for cross-linguistic comparison, diachronic analysis and natural language processing. The article, therefore, engages in four tasks: (i) identifying the Old English exponents of UD lexical categories; (ii) selecting the Old English exponents of UD morphological features; (iii) finding the areas of Old English morphology that require token indexing in the UD format; and (iv) checking on the relevance of the universal set of dependency relations. The data have been extracted from ParCorOEv2, an open access annotated parallel corpus Old English-English. The main conclusions are that the annotation format calls for two additional fields (gloss and morphological relatedness) and that enhanced dependencies are required in order to account for some syntactic phenomena.","PeriodicalId":42090,"journal":{"name":"Revista de Linguistica y Lenguas Aplicadas","volume":null,"pages":null},"PeriodicalIF":0.3000,"publicationDate":"2022-07-28","publicationTypes":"Journal Article","fieldsOfStudy":null,"isOpenAccess":false,"openAccessPdf":"","citationCount":"0","resultStr":null,"platform":"Semanticscholar","paperid":null,"PeriodicalName":"Revista de Linguistica y Lenguas Aplicadas","FirstCategoryId":"1085","ListUrlMain":"https://doi.org/10.4995/rlyla.2022.16787","RegionNum":0,"RegionCategory":null,"ArticlePicture":[],"TitleCN":null,"AbstractTextCN":null,"PMCID":null,"EPubDate":"","PubModel":"","JCR":"0","JCRName":"LANGUAGE & LINGUISTICS","Score":null,"Total":0}
引用次数: 0
Abstract
The aim of this article is to take the first steps toward the compilation of a treebank of Old English compatible with the framework of Universal Dependencies (UD). Such a treebank will comprise morphological and syntactic annotation of Old English texts adequate for cross-linguistic comparison, diachronic analysis and natural language processing. The article, therefore, engages in four tasks: (i) identifying the Old English exponents of UD lexical categories; (ii) selecting the Old English exponents of UD morphological features; (iii) finding the areas of Old English morphology that require token indexing in the UD format; and (iv) checking on the relevance of the universal set of dependency relations. The data have been extracted from ParCorOEv2, an open access annotated parallel corpus Old English-English. The main conclusions are that the annotation format calls for two additional fields (gloss and morphological relatedness) and that enhanced dependencies are required in order to account for some syntactic phenomena.
期刊介绍:
The Revista de Lingüística y Lenguas Aplicadas aims to contribute to thedissemination of scholarly research in the field of language study, especially thatof specialised languages. Whether from a theoretical or a practical perspective,contributions discussing any of the following areas are of particular interest: Discourse Analysis Language Teaching Terminology and Translation Languages for Specific Purposes (LSP) Computer-Assisted Language Learning (CALL) Its a peer-review yearly journal of linguistic studies, designed to target an international readership and to contribute to the promotion of knowledge regarding applied linguistics.