Combining WordNet and Word Embeddings in Data Augmentation for Legal Texts

Sezen Perçin, Andrea Galassi, F. Lagioia, Federico Ruggeri, Piera Santin, G. Sartor, Paolo Torroni
{"title":"Combining WordNet and Word Embeddings in Data Augmentation for Legal Texts","authors":"Sezen Perçin, Andrea Galassi, F. Lagioia, Federico Ruggeri, Piera Santin, G. Sartor, Paolo Torroni","doi":"10.18653/v1/2022.nllp-1.4","DOIUrl":null,"url":null,"abstract":"Creating balanced labeled textual corpora for complex tasks, like legal analysis, is a challenging and expensive process that often requires the collaboration of domain experts.To address this problem, we propose a data augmentation method based on the combination of GloVe word embeddings and the WordNet ontology.We present an example of application in the legal domain, specifically on decisions of the Court of Justice of the European Union.Our evaluation with human experts confirms that our method is more robust than the alternatives.","PeriodicalId":278495,"journal":{"name":"Proceedings of the Natural Legal Language Processing Workshop 2022","volume":"27 15 1","pages":"0"},"PeriodicalIF":0.0000,"publicationDate":"1900-01-01","publicationTypes":"Journal Article","fieldsOfStudy":null,"isOpenAccess":false,"openAccessPdf":"","citationCount":"3","resultStr":null,"platform":"Semanticscholar","paperid":null,"PeriodicalName":"Proceedings of the Natural Legal Language Processing Workshop 2022","FirstCategoryId":"1085","ListUrlMain":"https://doi.org/10.18653/v1/2022.nllp-1.4","RegionNum":0,"RegionCategory":null,"ArticlePicture":[],"TitleCN":null,"AbstractTextCN":null,"PMCID":null,"EPubDate":"","PubModel":"","JCR":"","JCRName":"","Score":null,"Total":0}
引用次数: 3

Abstract

Creating balanced labeled textual corpora for complex tasks, like legal analysis, is a challenging and expensive process that often requires the collaboration of domain experts.To address this problem, we propose a data augmentation method based on the combination of GloVe word embeddings and the WordNet ontology.We present an example of application in the legal domain, specifically on decisions of the Court of Justice of the European Union.Our evaluation with human experts confirms that our method is more robust than the alternatives.
查看原文
分享 分享
微信好友 朋友圈 QQ好友 复制链接
本刊更多论文
结合WordNet和词嵌入在法律文本数据增强中的应用
为复杂的任务(如法律分析)创建平衡的标记文本语料库是一个具有挑战性和昂贵的过程,通常需要领域专家的协作。为了解决这个问题,我们提出了一种基于GloVe词嵌入和WordNet本体相结合的数据增强方法。我们提出了一个在法律领域,特别是在欧洲联盟法院的判决中应用的例子。我们与人类专家的评估证实,我们的方法比替代方案更稳健。
本文章由计算机程序翻译,如有差异,请以英文原文为准。
求助全文
约1分钟内获得全文 去求助
来源期刊
自引率
0.00%
发文量
0
期刊最新文献
Detecting Relevant Differences Between Similar Legal Texts On What it Means to Pay Your Fair Share: Towards Automatically Mapping Different Conceptions of Tax Justice in Legal Research Literature Towards Cross-Domain Transferability of Text Generation Models for Legal Text Combining WordNet and Word Embeddings in Data Augmentation for Legal Texts
×
引用
GB/T 7714-2015
复制
MLA
复制
APA
复制
导出至
BibTeX EndNote RefMan NoteFirst NoteExpress
×
×
提示
您的信息不完整,为了账户安全,请先补充。
现在去补充
×
提示
您因"违规操作"
具体请查看互助需知
我知道了
×
提示
现在去查看 取消
×
提示
确定
0
微信
客服QQ
Book学术公众号 扫码关注我们
反馈
×
意见反馈
请填写您的意见或建议
请填写您的手机或邮箱
已复制链接
已复制链接
快去分享给好友吧!
我知道了
×
扫码分享
扫码分享
Book学术官方微信
Book学术文献互助
Book学术文献互助群
群 号:481959085
Book学术
文献互助 智能选刊 最新文献 互助须知 联系我们:info@booksci.cn
Book学术提供免费学术资源搜索服务,方便国内外学者检索中英文文献。致力于提供最便捷和优质的服务体验。
Copyright © 2023 Book学术 All rights reserved.
ghs 京公网安备 11010802042870号 京ICP备2023020795号-1