Toward the morpho-syntactic annotation of an Old English corpus with universal dependencies

IF 0.3 0 LANGUAGE & LINGUISTICS Revista de Linguistica y Lenguas Aplicadas Pub Date : 2022-07-28 DOI:10.4995/rlyla.2022.16787
Javier Martín Arista
{"title":"Toward the morpho-syntactic annotation of an Old English corpus with universal dependencies","authors":"Javier Martín Arista","doi":"10.4995/rlyla.2022.16787","DOIUrl":null,"url":null,"abstract":"The aim of this article is to take the first steps toward the compilation of a treebank of Old English compatible with the framework of Universal Dependencies (UD). Such a treebank will comprise morphological and syntactic annotation of Old English texts adequate for cross-linguistic comparison, diachronic analysis and natural language processing. The article, therefore, engages in four tasks: (i) identifying the Old English exponents of UD lexical categories; (ii) selecting the Old English exponents of UD morphological features; (iii) finding the areas of Old English morphology that require token indexing in the UD format; and (iv) checking on the relevance of the universal set of dependency relations. The data have been extracted from ParCorOEv2, an open access annotated parallel corpus Old English-English. The main conclusions are that the annotation format calls for two additional fields (gloss and morphological relatedness) and that enhanced dependencies are required in order to account for some syntactic phenomena.","PeriodicalId":42090,"journal":{"name":"Revista de Linguistica y Lenguas Aplicadas","volume":"1 1","pages":""},"PeriodicalIF":0.3000,"publicationDate":"2022-07-28","publicationTypes":"Journal Article","fieldsOfStudy":null,"isOpenAccess":false,"openAccessPdf":"","citationCount":"0","resultStr":null,"platform":"Semanticscholar","paperid":null,"PeriodicalName":"Revista de Linguistica y Lenguas Aplicadas","FirstCategoryId":"1085","ListUrlMain":"https://doi.org/10.4995/rlyla.2022.16787","RegionNum":0,"RegionCategory":null,"ArticlePicture":[],"TitleCN":null,"AbstractTextCN":null,"PMCID":null,"EPubDate":"","PubModel":"","JCR":"0","JCRName":"LANGUAGE & LINGUISTICS","Score":null,"Total":0}
引用次数: 0

Abstract

The aim of this article is to take the first steps toward the compilation of a treebank of Old English compatible with the framework of Universal Dependencies (UD). Such a treebank will comprise morphological and syntactic annotation of Old English texts adequate for cross-linguistic comparison, diachronic analysis and natural language processing. The article, therefore, engages in four tasks: (i) identifying the Old English exponents of UD lexical categories; (ii) selecting the Old English exponents of UD morphological features; (iii) finding the areas of Old English morphology that require token indexing in the UD format; and (iv) checking on the relevance of the universal set of dependency relations. The data have been extracted from ParCorOEv2, an open access annotated parallel corpus Old English-English. The main conclusions are that the annotation format calls for two additional fields (gloss and morphological relatedness) and that enhanced dependencies are required in order to account for some syntactic phenomena.
查看原文
分享 分享
微信好友 朋友圈 QQ好友 复制链接
本刊更多论文
具有普遍依赖关系的古英语语料库的词法句法注释
本文的目的是朝着编译与通用依赖关系(Universal Dependencies, UD)框架兼容的古英语树库迈出第一步。这样的树库将包括古英语文本的形态和句法注释,足以进行跨语言比较,历时分析和自然语言处理。因此,本文进行了四个方面的工作:(1)识别古英语中UD词汇类别的代表;(ii)选择UD形态特征的古英语指数;(iii)找出古英语词法中需要以UD格式进行标记索引的区域;(四)检验普遍依赖关系集的相关性。数据提取自ParCorOEv2,一个开放获取的注释并行语料库古英语-英语。主要结论是,注释格式需要两个额外的字段(gloss和形态学相关性),并且需要增强依赖性,以便解释一些语法现象。
本文章由计算机程序翻译,如有差异,请以英文原文为准。
求助全文
约1分钟内获得全文 去求助
来源期刊
CiteScore
0.50
自引率
25.00%
发文量
11
审稿时长
30 weeks
期刊介绍: The Revista de Lingüística y Lenguas Aplicadas aims to contribute to thedissemination of scholarly research in the field of language study, especially thatof specialised languages. Whether from a theoretical or a practical perspective,contributions discussing any of the following areas are of particular interest: Discourse Analysis Language Teaching Terminology and Translation Languages for Specific Purposes (LSP) Computer-Assisted Language Learning (CALL) Its a peer-review yearly journal of linguistic studies, designed to target an international readership and to contribute to the promotion of knowledge regarding applied linguistics.
期刊最新文献
Procesos de patrimonialización y creación de la identidad nacional El Aktionsart de pacientes con Alzhéimer: un análisis de corpus desde la Gramática del Papel y la Referencia Mapping the mental lexicon of EFL learners On experimental lexical production in Spanish as L1 and L2. Book review: Álvarez-Gil, F. J. (2022). Stance devices in tourism-related research articles: A corpus-based study
×
引用
GB/T 7714-2015
复制
MLA
复制
APA
复制
导出至
BibTeX EndNote RefMan NoteFirst NoteExpress
×
×
提示
您的信息不完整,为了账户安全,请先补充。
现在去补充
×
提示
您因"违规操作"
具体请查看互助需知
我知道了
×
提示
现在去查看 取消
×
提示
确定
0
微信
客服QQ
Book学术公众号 扫码关注我们
反馈
×
意见反馈
请填写您的意见或建议
请填写您的手机或邮箱
已复制链接
已复制链接
快去分享给好友吧!
我知道了
×
扫码分享
扫码分享
Book学术官方微信
Book学术文献互助
Book学术文献互助群
群 号:481959085
Book学术
文献互助 智能选刊 最新文献 互助须知 联系我们:info@booksci.cn
Book学术提供免费学术资源搜索服务,方便国内外学者检索中英文文献。致力于提供最便捷和优质的服务体验。
Copyright © 2023 Book学术 All rights reserved.
ghs 京公网安备 11010802042870号 京ICP备2023020795号-1