Improving the Lexical Ability of Pretrained Language Models for Unsupervised Neural Machine Translation

Unsupervised Neural Machine Translation with SMT as Posterior Regularization

Proceedings of the AAAI Conference on Artificial Intelligence ◽

10.1609/aaai.v33i01.3301241 ◽

2019 ◽

Vol 33 ◽

pp. 241-248 ◽

Cited By ~ 3

Author(s):

Shuo Ren ◽

Zhirui Zhang ◽

Shujie Liu ◽

Ming Zhou ◽

Shuai Ma

Keyword(s):

Machine Translation ◽

Language Models ◽

Translation Process ◽

Weak Supervision ◽

Neural Machine Translation ◽

Back Translation ◽

Negative Effect ◽

Model Training ◽

Cross Lingual ◽

Pseudo Data

Without real bilingual corpus available, unsupervised Neural Machine Translation (NMT) typically requires pseudo parallel data generated with the back-translation method for the model training. However, due to weak supervision, the pseudo data inevitably contain noises and errors that will be accumulated and reinforced in the subsequent training process, leading to bad translation performance. To address this issue, we introduce phrase based Statistic Machine Translation (SMT) models which are robust to noisy data, as posterior regularizations to guide the training of unsupervised NMT models in the iterative back-translation process. Our method starts from SMT models built with pre-trained language models and word-level translation tables inferred from cross-lingual embeddings. Then SMT and NMT models are optimized jointly and boost each other incrementally in a unified EM framework. In this way, (1) the negative effect caused by errors in the iterative back-translation process can be alleviated timely by SMT filtering noises from its phrase tables; meanwhile, (2) NMT can compensate for the deficiency of fluency inherent in SMT. Experiments conducted on en-fr and en-de translation tasks show that our method outperforms the strong baseline and achieves new state-of-the-art unsupervised machine translation performance.

Download Full-text

Towards Making the Most of BERT in Neural Machine Translation

Proceedings of the AAAI Conference on Artificial Intelligence ◽

10.1609/aaai.v34i05.6479 ◽

2020 ◽

Vol 34 (05) ◽

pp. 9378-9385

Author(s):

Jiacheng Yang ◽

Mingxuan Wang ◽

Hao Zhou ◽

Chengqi Zhao ◽

Weinan Zhang ◽

...

Keyword(s):

Machine Translation ◽

Language Processing ◽

State Of The Art ◽

Fine Tuning ◽

Language Models ◽

German Language ◽

Neural Machine Translation ◽

Dynamic Switching ◽

Previous State ◽

Language Pair

GPT-2 and BERT demonstrate the effectiveness of using pre-trained language models (LMs) on various natural language processing tasks. However, LM fine-tuning often suffers from catastrophic forgetting when applied to resource-rich tasks. In this work, we introduce a concerted training framework (CTnmt) that is the key to integrate the pre-trained LMs to neural machine translation (NMT). Our proposed CTnmt} consists of three techniques: a) asymptotic distillation to ensure that the NMT model can retain the previous pre-trained knowledge; b) a dynamic switching gate to avoid catastrophic forgetting of pre-trained knowledge; and c) a strategy to adjust the learning paces according to a scheduled policy. Our experiments in machine translation show CTnmt gains of up to 3 BLEU score on the WMT14 English-German language pair which even surpasses the previous state-of-the-art pre-training aided NMT by 1.4 BLEU score. While for the large WMT14 English-French task with 40 millions of sentence-pairs, our base model still significantly improves upon the state-of-the-art Transformer big model by more than 1 BLEU score.

Download Full-text

An Empirical Study on Learning Bug-Fixing Patches in the Wild via Neural Machine Translation

ACM Transactions on Software Engineering and Methodology ◽

10.1145/3340544 ◽

2019 ◽

Vol 28 (4) ◽

pp. 1-29 ◽

Cited By ~ 2

Author(s):

Michele Tufano ◽

Cody Watson ◽

Gabriele Bavota ◽

Massimiliano Di Penta ◽

Martin White ◽

...

Keyword(s):

Empirical Study ◽

Machine Translation ◽

Neural Machine Translation ◽

Bug Fixing ◽

In The Wild

Download Full-text

Neural Machine Translation for Semantic-Driven Q&A Systems in the Factory Planning

Procedia CIRP ◽

10.1016/j.procir.2021.01.044 ◽

2021 ◽

Vol 96 ◽

pp. 9-14

Author(s):

Uwe Dombrowski ◽

Alexander Reiswich ◽

Raphael Lamprecht

Keyword(s):

Machine Translation ◽

Neural Machine Translation ◽

Factory Planning

Download Full-text

A Neural Machine Translation Approach for Translating Malay Parliament Hansard to English Text

2020 International Conference on Asian Language Processing (IALP) ◽

10.1109/ialp51396.2020.9310470 ◽

2020 ◽

Author(s):

Yu-Zane Low ◽

Lay-Ki Soon ◽

Shageenderan Sapai

Keyword(s):

Machine Translation ◽

English Text ◽

Neural Machine Translation

Download Full-text

An Evaluation of Neural Machine Translation and Pre-trained Word Embeddings in Multilingual Neural Sentiment Analysis

2020 IEEE International Conference on Progress in Informatics and Computing (PIC) ◽

10.1109/pic50277.2020.9350849 ◽

2020 ◽

Author(s):

George Manias ◽

Argyro Mavrogiorgou ◽

Athanasios Kiourtis ◽

Dimosthenis Kyriazis

Keyword(s):

Sentiment Analysis ◽

Machine Translation ◽

Word Embeddings ◽

Neural Machine Translation

Download Full-text

Research on the Application of BERT in Mongolian-Chinese Neural Machine Translation

2021 13th International Conference on Machine Learning and Computing ◽

10.1145/3457682.3457744 ◽

2021 ◽

Author(s):

Xiu Zhi ◽

Siriguleng Wang

Keyword(s):

Machine Translation ◽

Neural Machine Translation

Download Full-text

Context- and sequence-aware convolutional recurrent encoder for neural machine translation

Proceedings of the 36th Annual ACM Symposium on Applied Computing ◽

10.1145/3412841.3442099 ◽

2021 ◽

Author(s):

Ritam Mallick ◽

Seba Susan ◽

Vaibhaw Agrawal ◽

Rizul Garg ◽

Prateek Rawal

Keyword(s):

Machine Translation ◽

Neural Machine Translation

Download Full-text

Linked Data Effectiveness in Neural Machine Translation

Proceedings of the 2020 4th International Symposium on Computer Science and Intelligent Control ◽

10.1145/3440084.3441214 ◽

2020 ◽

Author(s):

Benyamin Ahmadnia

Keyword(s):

Machine Translation ◽

Linked Data ◽

Neural Machine Translation

Download Full-text

Neural machine translation with a polysynthetic low resource language

Machine Translation ◽

10.1007/s10590-020-09255-9 ◽

2020 ◽

Vol 34 (4) ◽

pp. 325-346

Author(s):

John E. Ortega ◽

Richard Castro Mamani ◽

Kyunghyun Cho

Keyword(s):

Machine Translation ◽

Neural Machine Translation ◽

Low Resource

Download Full-text