结合WordNet和词嵌入在法律文本数据增强中的应用

Proceedings of the Natural Legal Language Processing Workshop 2022 Pub Date : 1900-01-01 DOI:10.18653/v1/2022.nllp-1.4

Sezen Perçin, Andrea Galassi, F. Lagioia, Federico Ruggeri, Piera Santin, G. Sartor, Paolo Torroni

{"title":"结合WordNet和词嵌入在法律文本数据增强中的应用","authors":"Sezen Perçin, Andrea Galassi, F. Lagioia, Federico Ruggeri, Piera Santin, G. Sartor, Paolo Torroni","doi":"10.18653/v1/2022.nllp-1.4","DOIUrl":null,"url":null,"abstract":"Creating balanced labeled textual corpora for complex tasks, like legal analysis, is a challenging and expensive process that often requires the collaboration of domain experts.To address this problem, we propose a data augmentation method based on the combination of GloVe word embeddings and the WordNet ontology.We present an example of application in the legal domain, specifically on decisions of the Court of Justice of the European Union.Our evaluation with human experts confirms that our method is more robust than the alternatives.","PeriodicalId":278495,"journal":{"name":"Proceedings of the Natural Legal Language Processing Workshop 2022","volume":"27 15 1","pages":"0"},"PeriodicalIF":0.0000,"publicationDate":"1900-01-01","publicationTypes":"Journal Article","fieldsOfStudy":null,"isOpenAccess":false,"openAccessPdf":"","citationCount":"3","resultStr":"{\"title\":\"Combining WordNet and Word Embeddings in Data Augmentation for Legal Texts\",\"authors\":\"Sezen Perçin, Andrea Galassi, F. Lagioia, Federico Ruggeri, Piera Santin, G. Sartor, Paolo Torroni\",\"doi\":\"10.18653/v1/2022.nllp-1.4\",\"DOIUrl\":null,\"url\":null,\"abstract\":\"Creating balanced labeled textual corpora for complex tasks, like legal analysis, is a challenging and expensive process that often requires the collaboration of domain experts.To address this problem, we propose a data augmentation method based on the combination of GloVe word embeddings and the WordNet ontology.We present an example of application in the legal domain, specifically on decisions of the Court of Justice of the European Union.Our evaluation with human experts confirms that our method is more robust than the alternatives.\",\"PeriodicalId\":278495,\"journal\":{\"name\":\"Proceedings of the Natural Legal Language Processing Workshop 2022\",\"volume\":\"27 15 1\",\"pages\":\"0\"},\"PeriodicalIF\":0.0000,\"publicationDate\":\"1900-01-01\",\"publicationTypes\":\"Journal Article\",\"fieldsOfStudy\":null,\"isOpenAccess\":false,\"openAccessPdf\":\"\",\"citationCount\":\"3\",\"resultStr\":null,\"platform\":\"Semanticscholar\",\"paperid\":null,\"PeriodicalName\":\"Proceedings of the Natural Legal Language Processing Workshop 2022\",\"FirstCategoryId\":\"1085\",\"ListUrlMain\":\"https://doi.org/10.18653/v1/2022.nllp-1.4\",\"RegionNum\":0,\"RegionCategory\":null,\"ArticlePicture\":[],\"TitleCN\":null,\"AbstractTextCN\":null,\"PMCID\":null,\"EPubDate\":\"\",\"PubModel\":\"\",\"JCR\":\"\",\"JCRName\":\"\",\"Score\":null,\"Total\":0}","platform":"Semanticscholar","paperid":null,"PeriodicalName":"Proceedings of the Natural Legal Language Processing Workshop 2022","FirstCategoryId":"1085","ListUrlMain":"https://doi.org/10.18653/v1/2022.nllp-1.4","RegionNum":0,"RegionCategory":null,"ArticlePicture":[],"TitleCN":null,"AbstractTextCN":null,"PMCID":null,"EPubDate":"","PubModel":"","JCR":"","JCRName":"","Score":null,"Total":0}

引用次数: 3

摘要

为复杂的任务(如法律分析)创建平衡的标记文本语料库是一个具有挑战性和昂贵的过程，通常需要领域专家的协作。为了解决这个问题，我们提出了一种基于GloVe词嵌入和WordNet本体相结合的数据增强方法。我们提出了一个在法律领域，特别是在欧洲联盟法院的判决中应用的例子。我们与人类专家的评估证实，我们的方法比替代方案更稳健。

本文章由计算机程序翻译，如有差异，请以英文原文为准。

查看原文本刊更多论文

Combining WordNet and Word Embeddings in Data Augmentation for Legal Texts

Creating balanced labeled textual corpora for complex tasks, like legal analysis, is a challenging and expensive process that often requires the collaboration of domain experts.To address this problem, we propose a data augmentation method based on the combination of GloVe word embeddings and the WordNet ontology.We present an example of application in the legal domain, specifically on decisions of the Court of Justice of the European Union.Our evaluation with human experts confirms that our method is more robust than the alternatives.

求助全文

通过发布文献求助，成功后即可免费获取论文全文。去求助

来源期刊

Proceedings of the Natural Legal Language Processing Workshop 2022

自引率

0.00%

发文量