Domain Adaptation and Multi-Domain Adaptation for Neural Machine Translation: A Survey

Danielle Saunders

doi:10.1613/jair.1.13566

PDF

Published: Sep 29, 2022

DOI: https://doi.org/10.1613/jair.1.13566

Keywords:

machine translation, natural language

Danielle Saunders

a:1:{s:5:"en_US";s:7:"SDL plc";}

Abstract

The development of deep learning techniques has allowed Neural Machine Translation (NMT) models to become extremely powerful, given sufficient training data and training time. However, systems struggle when translating text from a new domain with a distinct style or vocabulary. Fine-tuning on in-domain data allows good domain adaptation, but requires sufficient relevant bilingual data. Even if this is available, simple fine-tuning can cause overfitting to new data and catastrophic forgetting of previously learned behaviour.

We survey approaches to domain adaptation for NMT, particularly where a system may need to translate across multiple domains. We divide techniques into those revolving around data selection or generation, model architecture, parameter adaptation procedure, and inference procedure. We finally highlight the benefits of domain adaptation and multidomain adaptation techniques to other lines of NMT research.

Issue

Vol. 75 (2022)

Section

Articles

Article Sidebar

Main Article Content

Abstract

Article Details