Improving Neural Machine Translation Through Code-Mixed Data Augmentation

IF 1.7 4区计算机科学 Q3 COMPUTER SCIENCE, ARTIFICIAL INTELLIGENCE Computational Intelligence Pub Date : 2025-03-06 DOI:10.1111/coin.70033

Ramakrishna Appicharla, Kamal Kumar Gupta, Asif Ekbal, Pushpak Bhattacharyya

引用次数: 0

Abstract

This paper studies neural machine translation (NMT) of code-mixed (CM) text. Specifically, we generate synthetic CM data and how it can be used to improve the translation performance of NMT through the data augmentation strategy. We conduct experiments on three data augmentation approaches viz. CM-Augmentation, CM-Concatenation, and Multi-Encoder approaches, and the latter two approaches are inspired by document-level NMT, where we use synthetic CM data as context to improve the performance of the NMT models. We conduct experiments on three language pairs, viz. Hindi–English, Telugu–English and Czech–English. Experimental results demonstrate that the proposed approaches significantly improve performance over the baseline model trained without data augmentation and over the existing data augmentation strategies. The CM-Concatenation model attains the best performance.

查看原文

微信好友朋友圈 QQ好友复制链接

本刊更多论文

通过代码混合数据增强改进神经机器翻译

本文对混合码文本的神经机器翻译进行了研究。具体来说，我们生成了合成CM数据，以及如何通过数据增强策略来提高NMT的翻译性能。我们对三种数据增强方法进行了实验，即CM- augmentation， CM- concatation和Multi-Encoder方法，后两种方法受到文档级NMT的启发，其中我们使用合成CM数据作为上下文来提高NMT模型的性能。我们对三种语言进行了实验，即印地语-英语、泰卢语-英语和捷克语-英语。实验结果表明，与没有数据增强训练的基线模型和现有的数据增强策略相比，所提出的方法显著提高了性能。cm - concatation模型的性能最好。

本文章由计算机程序翻译，如有差异，请以英文原文为准。

求助全文

约1分钟内获得全文去求助

来源期刊

Computational Intelligence 工程技术-计算机：人工智能

CiteScore

6.90

自引率

3.60%

发文量

审稿时长

>12 weeks

期刊介绍： This leading international journal promotes and stimulates research in the field of artificial intelligence (AI). Covering a wide range of issues - from the tools and languages of AI to its philosophical implications - Computational Intelligence provides a vigorous forum for the publication of both experimental and theoretical research, as well as surveys and impact studies. The journal is designed to meet the needs of a wide range of AI workers in academic and industrial research.