机器翻译中情感偏差的测量

International Conference on Text, Speech and Dialogue Pub Date : 2023-06-12 DOI:10.48550/arXiv.2306.07152

Kai Hartung, Aaricia Herygers, Shubham Kurlekar, Khabbab Zakaria, Taylan Volkan, Sören Gröttrup, Munir Georges

{"title":"机器翻译中情感偏差的测量","authors":"Kai Hartung, Aaricia Herygers, Shubham Kurlekar, Khabbab Zakaria, Taylan Volkan, Sören Gröttrup, Munir Georges","doi":"10.48550/arXiv.2306.07152","DOIUrl":null,"url":null,"abstract":"Biases induced to text by generative models have become an increasingly large topic in recent years. In this paper we explore how machine translation might introduce a bias in sentiments as classified by sentiment analysis models. For this, we compare three open access machine translation models for five different languages on two parallel corpora to test if the translation process causes a shift in sentiment classes recognized in the texts. Though our statistic test indicate shifts in the label probability distributions, we find none that appears consistent enough to assume a bias induced by the translation process.","PeriodicalId":358274,"journal":{"name":"International Conference on Text, Speech and Dialogue","volume":"40 1","pages":"0"},"PeriodicalIF":0.0000,"publicationDate":"2023-06-12","publicationTypes":"Journal Article","fieldsOfStudy":null,"isOpenAccess":false,"openAccessPdf":"","citationCount":"0","resultStr":"{\"title\":\"Measuring Sentiment Bias in Machine Translation\",\"authors\":\"Kai Hartung, Aaricia Herygers, Shubham Kurlekar, Khabbab Zakaria, Taylan Volkan, Sören Gröttrup, Munir Georges\",\"doi\":\"10.48550/arXiv.2306.07152\",\"DOIUrl\":null,\"url\":null,\"abstract\":\"Biases induced to text by generative models have become an increasingly large topic in recent years. In this paper we explore how machine translation might introduce a bias in sentiments as classified by sentiment analysis models. For this, we compare three open access machine translation models for five different languages on two parallel corpora to test if the translation process causes a shift in sentiment classes recognized in the texts. Though our statistic test indicate shifts in the label probability distributions, we find none that appears consistent enough to assume a bias induced by the translation process.\",\"PeriodicalId\":358274,\"journal\":{\"name\":\"International Conference on Text, Speech and Dialogue\",\"volume\":\"40 1\",\"pages\":\"0\"},\"PeriodicalIF\":0.0000,\"publicationDate\":\"2023-06-12\",\"publicationTypes\":\"Journal Article\",\"fieldsOfStudy\":null,\"isOpenAccess\":false,\"openAccessPdf\":\"\",\"citationCount\":\"0\",\"resultStr\":null,\"platform\":\"Semanticscholar\",\"paperid\":null,\"PeriodicalName\":\"International Conference on Text, Speech and Dialogue\",\"FirstCategoryId\":\"1085\",\"ListUrlMain\":\"https://doi.org/10.48550/arXiv.2306.07152\",\"RegionNum\":0,\"RegionCategory\":null,\"ArticlePicture\":[],\"TitleCN\":null,\"AbstractTextCN\":null,\"PMCID\":null,\"EPubDate\":\"\",\"PubModel\":\"\",\"JCR\":\"\",\"JCRName\":\"\",\"Score\":null,\"Total\":0}","platform":"Semanticscholar","paperid":null,"PeriodicalName":"International Conference on Text, Speech and Dialogue","FirstCategoryId":"1085","ListUrlMain":"https://doi.org/10.48550/arXiv.2306.07152","RegionNum":0,"RegionCategory":null,"ArticlePicture":[],"TitleCN":null,"AbstractTextCN":null,"PMCID":null,"EPubDate":"","PubModel":"","JCR":"","JCRName":"","Score":null,"Total":0}

引用次数: 0

摘要

近年来，生成模型对文本产生的偏见已经成为一个越来越大的话题。在本文中，我们探讨了机器翻译如何在情感分析模型分类的情感中引入偏见。为此，我们在两个平行语料库上比较了五种不同语言的三种开放存取机器翻译模型，以测试翻译过程是否会导致文本中识别的情感类别发生变化。虽然我们的统计检验表明标签概率分布发生了变化，但我们发现没有一个数据看起来足够一致，足以假设翻译过程引起的偏差。

本文章由计算机程序翻译，如有差异，请以英文原文为准。

查看原文

微信好友朋友圈 QQ好友复制链接

本刊更多论文

Measuring Sentiment Bias in Machine Translation

Biases induced to text by generative models have become an increasingly large topic in recent years. In this paper we explore how machine translation might introduce a bias in sentiments as classified by sentiment analysis models. For this, we compare three open access machine translation models for five different languages on two parallel corpora to test if the translation process causes a shift in sentiment classes recognized in the texts. Though our statistic test indicate shifts in the label probability distributions, we find none that appears consistent enough to assume a bias induced by the translation process.

求助全文

通过发布文献求助，成功后即可免费获取论文全文。去求助

来源期刊

International Conference on Text, Speech and Dialogue

自引率

0.00%

发文量

期刊最新文献

Advancing Hungarian Text Processing with HuSpaCy: Efficient and Accurate NLP Pipelines A Dataset and Strong Baselines for Classification of Czech News Texts Measuring Sentiment Bias in Machine Translation Transfer Learning of Transformer-based Speech Recognition Models from Czech to Slovak Sub 8-Bit Quantization of Streaming Keyword Spotting Models for Embedded Chipsets